Skip to content
@Relaxed-System-Lab

Relaxed System Lab

The research lab lead by Binhang Yuan @ CSE HKUST

Popular repositories Loading

  1. Flash-Sparse-Attention Flash-Sparse-Attention Public

    🚀🚀 Efficient implementations of Native Sparse Attention

    Python 603 15

  2. multi-actor-data-selection multi-actor-data-selection Public

    This is the repo for the paper Multi-Agent Collaborative Data Selection for Efficient LLM Pretraining.

    Python 49 3

  3. HexGen HexGen Public

    [ICML 2024] Serving LLMs on heterogeneous decentralized clusters.

    Python 38 3

  4. HexiScale HexiScale Public

    Accommodating Large Language Model Training over Heterogeneous Environment.

    Python 36 10

  5. HKUST-COMP6211J-2025fall HKUST-COMP6211J-2025fall Public

    24 1

  6. UltraLLaDA UltraLLaDA Public

    We introduce UltraLLaDA , a scaled variant of LLaDA-8B-Base that extends the context length up to 128K tokens with light-weight post-training, enabling long-context comprehension and generation.

    Python 15

Repositories

Showing 10 of 24 repositories
  • Flash-Sparse-Attention Public

    🚀🚀 Efficient implementations of Native Sparse Attention

    Relaxed-System-Lab/Flash-Sparse-Attention's past year of commit activity
    Python 603 Apache-2.0 15 7 0 Updated Sep 17, 2026
  • KernelBraid Public

    Prompt-driven optimization harness and bitwise-aligned attention kernels for RL post-training.

    Relaxed-System-Lab/KernelBraid's past year of commit activity
    Cuda 2 Apache-2.0 0 0 0 Updated Sep 17, 2026
  • sglang Public Forked from sgl-project/sglang

    SGLang is a high-performance serving framework for large language models and multimodal models.

    Relaxed-System-Lab/sglang's past year of commit activity
    Python 0 Apache-2.0 9,034 0 0 Updated Aug 19, 2026
  • ClimateAgent Public
    Relaxed-System-Lab/ClimateAgent's past year of commit activity
    Python 3 Apache-2.0 0 1 0 Updated Aug 10, 2026
  • HexGen-3 Public

    [ICML 2024, ICLR 2025, ICML 2025, ICML 2026] HexGen: LLM Serving over Heterogeneous GPUs; HexGen-2: PD Disaggregation over Heterogeneous GPUs; Demystifying Cost-Efficiency of LLM Serving over Heterogeneous GPUs; HexGen-3: Fully Disaggregated Serving and Resource Autoscaling over Heterogeneous GPUs.

    Relaxed-System-Lab/HexGen-3's past year of commit activity
    Python 4 2 0 0 Updated Jul 30, 2026
  • Relaxed-System-Lab/HKUST-COMP4551-2026spring's past year of commit activity
    15 0 0 0 Updated May 7, 2026
  • V3DB Public

    IVF-PQ zero-knowledge prove system.

    Relaxed-System-Lab/V3DB's past year of commit activity
    Python 6 1 0 0 Updated May 1, 2026
  • Text2SQL-CHESS Public Forked from ShayanTalaei/CHESS

    A Text2SQL Serving System based on CHESS Framework

    Relaxed-System-Lab/Text2SQL-CHESS's past year of commit activity
    Python 0 Apache-2.0 84 0 0 Updated Mar 20, 2026
  • pypto-lib Public Forked from hw-native-sys/pypto-lib
    Relaxed-System-Lab/pypto-lib's past year of commit activity
    Python 0 64 0 0 Updated Mar 16, 2026
  • ray Public Forked from ray-project/ray

    Ray is an AI compute engine. Ray consists of a core distributed runtime and a set of AI Libraries for accelerating ML workloads.

    Relaxed-System-Lab/ray's past year of commit activity
    Python 0 Apache-2.0 8,262 0 0 Updated Feb 13, 2026

People

This organization has no public members. You must be a member to see who’s a part of this organization.

Most used topics

Loading…