There is a huge gap in open source AI accelerators, so I implemented mine. Popular and well known ones are already legacy and doesn't support contemporary operations like Attention. Here is what makes mine special: Attention mechanism smelted directly into silicon Prototyped end-to-end on FPGA (AWS