Skip to content
View Aaronhuang-778's full-sized avatar

Organizations

@Efficient-Large-Model

Block or report Aaronhuang-778

Block user

Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.

You must be logged in to block users.

Content in all repositories owned by your account will be closed.
Maximum 250 characters. Please don’t include any personal information such as legal names or email addresses. Markdown is supported. This note will only be visible to you.
Report abuse

Contact GitHub support about this user’s behavior. Learn more about reporting abuse.

Report abuse

Pinned Loading

  1. NVlabs/LongLive NVlabs/LongLive Public

    Long Video Gen Infrastructure

    Python 2.6k 248

  2. NVlabs/Long-RL NVlabs/Long-RL Public

    Long-RL: Scaling RL to Long Sequences (NeurIPS 2025)

    Python 728 31

  3. NVlabs/QeRL NVlabs/QeRL Public

    [ICLR 2026]QeRL enables RL for 32B LLMs on a single H100 GPU.

    Python 516 54

  4. WeianMao/triattention WeianMao/triattention Public

    TriAttention — Efficient long reasoning with trigonometric KV cache compression. Enables OpenClaw local deployment on memory-constrained GPUs.

    Python 846 79

  5. BiLLM BiLLM Public

    [ICML 2024] BiLLM: Pushing the Limit of Post-Training Quantization for LLMs

    Python 235 18

  6. Mixture-Compressor-MoE Mixture-Compressor-MoE Public

    [ICLR 2025, IEEE TPAMI 2026] Mixture Compressor & MC#

    Python 77 9