- Toronto, Ontario, Canada
- https://shawsilicon.ai
Popular repositories Loading
-
fpga-ai-accelerator
fpga-ai-accelerator PublicOpen-source 8×8 INT8 systolic array inference accelerator — from RTL to bitstream
-
kvcache-compress-engine
kvcache-compress-engine PublicCOMPRESSION_RATIO = 8× — Hardware KV-cache compression engine eliminates the memory wall in LLM inference. 1,155 LUTs, 1 DSP, 400 MHz.
-
flashattn-softmax-engine
flashattn-softmax-engine PublicPERF_STALL_CYCLES = 0 — Hardened softmax pipeline eliminates the #1 bottleneck in transformer attention. 550 LUTs, 0 DSPs, 400 MHz.
-
gnss_spoof_jam_detector_hls_rtl
gnss_spoof_jam_detector_hls_rtl PublicStreaming GNSS spoofing/jamming detection accelerator - HLS metric engine + RTL NCO/PRN/AXI streaming, verified under cycle-level backpressure in XSim and Zynq UltraScale MPSoC. Vitis HLS 2025.2 sy…
SystemVerilog 4
-
rope-rotary-engine
rope-rotary-engine PublicOpen-source SystemVerilog hardware accelerator for Rotary Position Embedding (RoPE) - the positional encoding used by Llama 3/4, DeepSeek V3, Qwen 2/3.
If the problem persists, check the GitHub status page or contact support.
