Skip to content
View Aaronhuang-778's full-sized avatar

Organizations

@Efficient-Large-Model

Block or report Aaronhuang-778

Block user

Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.

You must be logged in to block users.

Maximum 250 characters. Please don't include any personal information such as legal names or email addresses. Markdown supported. This note will be visible to only you.
Report abuse

Contact GitHub support about this user’s behavior. Learn more about reporting abuse.

Report abuse

Pinned Loading

  1. NVlabs/Long-RL NVlabs/Long-RL Public

    Long-RL: Scaling RL to Long Sequences (NeurIPS 2025)

    Python 647 23

  2. NVlabs/QeRL NVlabs/QeRL Public

    QeRL enables RL for 32B LLMs on a single H100 GPU.

    Python 382 29

  3. BiLLM BiLLM Public

    [ICML 2024] BiLLM: Pushing the Limit of Post-Training Quantization for LLMs

    Python 228 16

  4. Mixture-Compressor-MoE Mixture-Compressor-MoE Public

    [ICLR 2025] Mixture Compressor for Mixture-of-Experts LLMs Gains More

    Python 60 3

  5. SliM-LLM SliM-LLM Public

    [ICML 2025] SliM-LLM: Salience-Driven Mixed-Precision Quantization for Large Language Models

    Python 45 3

  6. hshjerry/VideoEspresso hshjerry/VideoEspresso Public

    [CVPR 2025 Oral] VideoEspresso: A Large-Scale Chain-of-Thought Dataset for Fine-Grained Video Reasoning via Core Frame Selection

    Python 125 4