Our big difference that you were trying to remember was our use of scratchpad memory which is simply stated as this: We can radically reduce power consumption, increase density, and increase speed of on chip memory (SRAM) by removing the traditional hardware caching system. We instead use a purely software managed memory system, through some very fancy (or do I dare say "smart") compiler techniques that are enabled by a a very simplified architecture and the ability to guarantee latency for all memory operations. We can still have main system memory (meaning DRAM), it is just instead of having a bunch of complex hardware that burns a lot of power and wastes a lot of space in order to automatically fetch pages out of DRAM, we structure the code for each core to efficiently pull only the data necessary when it is needed.
One of the things I wondered about is how Linux is adapted to such an architecture. I couldn't find anything from REX on operating system support.
As for REX, as dnautics said, we've been mostly focused on running raw compute kernels on the current simulated versions (software and FPGA), and for the soon to be in hand silicon (coming this fall)... one of our projects internally is to port the L4 microkernel, and a telecom focused RTOS, but that is as far as our operating systems plans go for the near future. I'd also love to get a Plan 9/inferno demo running on it for fun, but we've got more important work to do at the moment.
The point about the OS related to the Linux-based one for Sunway, but maybe it only runs on the management processor anyhow, with just offload to the others. I'd commented in that respect that we really don't want something like Linux in an ideal world, so I'm pleased to see mention of L4.
Thanks for the comments, and good luck.
Videos:
https://www.youtube.com/channel/UCKdGg6hZoUYnjyRUb08Kjbg/
The Mill CPU Architecture – The Compiler [video] (youtube.com):
https://news.ycombinator.com/item?id=9856334
Wiki:
Maybe a Massively Parallel Processor Array thing? https://en.wikipedia.org/wiki/Massively_parallel_processor_a...
For future reference, here is a link talking about the arch mentioned so that no one else has to wade through google's results for "fleet architecture":
http://arc.cecs.pdx.edu/publications
edit:
previous (2009) discussion about fleet: https://news.ycombinator.com/item?id=723882