Skip to content
View vokkko's full-sized avatar

Block or report vokkko

Report abuse

Contact GitHub support about this user’s behavior. Learn more about reporting abuse.

Report abuse

Popular repositories Loading

  1. Awesome-Efficient-LLM Awesome-Efficient-LLM Public

    Forked from horseee/Awesome-Efficient-LLM

    A curated list for Efficient Large Language Models

    Python

  2. auto-round auto-round Public

    Forked from intel/auto-round

    SOTA Weight-only Quantization Algorithm for LLMs. This is official implementation of "Optimize Weight Rounding via Signed Gradient Descent for the Quantization of LLMs"

    Python

  3. EfficientDM EfficientDM Public

    Forked from ThisisBillhe/EfficientDM

    [ICLR 2024 Spotlight] This is the official PyTorch implementation of "EfficientDM: Efficient Quantization-Aware Fine-Tuning of Low-Bit Diffusion Models"

    Jupyter Notebook

  4. Quest Quest Public

    Forked from mit-han-lab/Quest

    [ICML 2024] Quest: Query-Aware Sparsity for Efficient Long-Context LLM Inference

    Cuda

  5. llmc llmc Public

    Forked from ModelTC/LightCompress

    This is the official PyTorch implementation of "LLMC: Benchmarking Large Language Model Quantization with a Versatile Compression Toolkit".

    Python

  6. bilivideos bilivideos Public

    Forked from cauyxy/bilivideos

    Jupyter Notebook