GateGPT: 56k tokens per second Transformer (KV cache) on FPGA at 80 MHz

(twitter.com)

32 points | by laxmena 3 hours ago ago

10 comments