Skip to content
View htqin's full-sized avatar
🦊
Focusing
🦊
Focusing

Highlights

  • Pro

Block or report htqin

Block user

Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.

You must be logged in to block users.

Content in all repositories owned by your account will be closed.
Maximum 250 characters. Please don’t include any personal information such as legal names or email addresses. Markdown is supported. This note will only be visible to you.
Report abuse

Contact GitHub support about this user’s behavior. Learn more about reporting abuse.

Report abuse

Pinned Loading

  1. AI-Efficiency/Awesome-Model-Quantization AI-Efficiency/Awesome-Model-Quantization Public

    A curated collection of papers, benchmarks, surveys, and tools for model quantization, covering low-bit networks, LLMs, multimodal and generative models, vector and lattice quantization, and effici…

    2.4k 242

  2. AI-Efficiency/Awesome-Efficient-Diffusion AI-Efficiency/Awesome-Efficient-Diffusion Public

    A curated list of papers and code on efficient diffusion models for image, video, world modeling, and language generation. Covering acceleration, quantization, compression, caching, and distillation.

    208 8

  3. AI-Efficiency/BiBERT AI-Efficiency/BiBERT Public

    This project is the official implementation of our accepted ICLR 2022 paper BiBERT: Accurate Fully Binarized BERT.

    Python 89 6

  4. AI-Efficiency/IR-QLoRA AI-Efficiency/IR-QLoRA Public

    [ICML 2024 Oral] This project is the official implementation of our Accurate LoRA-Finetuning Quantization of LLMs via Information Retention

    Python 65 4

  5. AI-Efficiency/BiBench AI-Efficiency/BiBench Public

    [ICML 2023] This project is the official implementation of our accepted ICML 2023 paper BiBench: Benchmarking and Analyzing Network Binarization.

    Python 55 5

  6. IR-Net IR-Net Public

    [CVPR 2020] This project is the PyTorch implementation of our accepted CVPR 2020 paper : forward and backward information retention for accurate binary neural networks.

    Python 181 34