
apex-inference-chip
SigmanticAI
An inference chip design that runs a real LLM (Qwen2.5-0.5B) on FPGA — one transformer decoder layer in RTL, every silicon value bit-exact against a golden model. 0.56 tok/s measured, a 140× climb, full evidence trail.
Python
Apache License 2.0680
Stars
1
Forks
364
Watchers
0
Issues
Star 增长
今日+2
近 7 天+138
近 30 天+138
综合评分74.9
默认分支main
暂无 README 内容
项目可能尚未同步完成,请稍后查看