Calyx, a Compiler Infrastructure for Accelerator Generators
calyxir.org
calyxir.org
Bias: I implemented an integration for iree.
If you work on FPGA’s, are there any HLS systems that you think are good enough to implement most of the techniques we see in the AI research papers? Especially those reducing the cost of pre-training or increasing the context of models.
you're talking about two different worlds that are universes apart.
fact 1: no matter what anyone tells you (calyx included), you cannot splat NNs onto silicon (neither FPGA nor ASIC) - i spent probably thousands of hours trying to do it for CNNs and the results were completely unimpressive. why? because routing and statically scheduling is NP-hard. what people do (what i'm doing now) is build GEMM accelerators and then contort their NN layers into sequences of GEMMs.
fact 2: HLS will never work for anything larger than toy problems because scheduling is NP-hard (see 1).
> I used to look for tools that bypass most of the hardware design
so you can absolutely, without a doubt, give up on this dream (until knuth proves NP=P).