Skip to content
View taitashaw's full-sized avatar

Block or report taitashaw

Block user

Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.

You must be logged in to block users.

Maximum 250 characters. Please don’t include any personal information such as legal names or email addresses. Markdown is supported. This note will only be visible to you.
Report abuse

Contact GitHub support about this user’s behavior. Learn more about reporting abuse.

Report abuse

Popular repositories Loading

  1. fpga-ai-accelerator fpga-ai-accelerator Public

    Open-source 8×8 INT8 systolic array inference accelerator — from RTL to bitstream

    SystemVerilog 19 3

  2. HLS_FPGA HLS_FPGA Public

    🚀 Real-Time High-Level Synthesis (HLS) Projects for UltraScale+ FPGAs. Accelerated design pipelines, dataflow architectures, and low-latency compute cores — powered by Vitis HLS, AXI4-Stream, and s…

    C++ 15 5

  3. kvcache-compress-engine kvcache-compress-engine Public

    COMPRESSION_RATIO = 8× — Hardware KV-cache compression engine eliminates the memory wall in LLM inference. 1,155 LUTs, 1 DSP, 400 MHz.

    SystemVerilog 5 2

  4. flashattn-softmax-engine flashattn-softmax-engine Public

    PERF_STALL_CYCLES = 0 — Hardened softmax pipeline eliminates the #1 bottleneck in transformer attention. 550 LUTs, 0 DSPs, 400 MHz.

    SystemVerilog 4 1

  5. rope-hls-vs-rtl rope-hls-vs-rtl Public

    Rotary Position Embedding in hardware: Vitis HLS versus hand-written SystemVerilog RTL, compared on resources, latency and timing closure.

    SystemVerilog 3

  6. gnss_spoof_jam_detector_hls_rtl gnss_spoof_jam_detector_hls_rtl Public

    Streaming GNSS spoofing/jamming detection accelerator - HLS metric engine + RTL NCO/PRN/AXI streaming, verified under cycle-level backpressure in XSim and Zynq UltraScale MPSoC. Vitis HLS 2025.2 sy…

    SystemVerilog 3