https://github.com/SigmanticAI/apex-inference-chip
FPGA实现Transformer推理芯片
在 FPGA 上运行真实 LLM 的推理芯片设计,速度 0.56 tok/s。
- 首次收录:
- 2026-08-20 08:00
- 出现过:
- 2026/08/23 08:00 · 2026/08/22 12:00 · 2026/08/22 08:00 · 2026/08/20 12:00 · 2026/08/20 08:00
https://github.com/SigmanticAI/apex-inference-chip
在 FPGA 上运行真实 LLM 的推理芯片设计,速度 0.56 tok/s。