compression

20 progetti condividono questo topic GitHub

compression — DeepSpeed ★42.7kcompressionPaddleNLP — ★13kSimpleMem — ★3.7kvortex — ★3.1kaimet — ★2.7kbolt — ★2.5kAwesome-Efficient-LLM — ★2kCompressAI — ★1.6kmodel-optimization — ★1.6kAwesome-Knowledge-Distillation-of-LLMs — ★1.3knncf — ★1.2kMindPipe — ★1kswin2sr — ★689DFloat11 — ★642wavemap — ★564KVQuant — ★431Context-Engine — ★400BK-SDM — ★319picollm — ★314distill — ★174PaddleNLP★ 13kSimpleMem★ 3.7kvortex★ 3.1kaimet★ 2.7kbolt★ 2.5kAwesome-Efficient-LLM★ 2kCompressAI★ 1.6kmodel-optimization★ 1.6kAwesome-Knowledge-Distil…★ 1.3knncf★ 1.2kMindPipe★ 1kswin2sr★ 689DFloat11★ 642wavemap★ 564KVQuant★ 431Context-Engine★ 400BK-SDM★ 319picollm★ 314distill★ 174

Le linee collegano membri che sono misurabilmente correlati tra loro. La dimensione dei punti riflette le stelle.

🧬 Membri
DeepSpeed
DeepSpeed is a deep learning optimization library that makes distributed training and inference easy,…
★ 42.7k
PaddleNLP
Easy-to-use and powerful LLM and SLM library with awesome model zoo.
★ 13k
SimpleMem
SimpleMem: Efficient Lifelong Memory for LLM Agents — Text & Multimodal
★ 3.7k
vortex
An extensible, state-of-the-art framework for columnar compression, and the fastest FOSS columnar file…
★ 3.1k
aimet
AIMET is a library that provides advanced quantization and compression techniques for trained neural network…
★ 2.7k
bolt
10x faster matrix and vector operations
★ 2.5k
Awesome-Efficient-LLM
A curated list for Efficient Large Language Models
★ 2k
CompressAI
A PyTorch library and evaluation platform for end-to-end compression research
★ 1.6k
model-optimization
A toolkit to optimize ML models for deployment for Keras and TensorFlow, including quantization and pruning.
★ 1.6k
Awesome-Knowledge-Distillation-of-LLMs
This repository collects papers for "A Survey on Knowledge Distillation of Large Language Models". We break…
★ 1.3k
nncf
Neural Network Compression Framework for enhanced OpenVINO™ inference
★ 1.2k
MindPipe
A powerful model compression framework for LLMs and LVLMs, adapted for NVIDIA GPUs and Huawei Ascend NPUs.
★ 1k
swin2sr
[ECCV] Swin2SR: SwinV2 Transformer for Compressed Image Super-Resolution and Restoration. Advances in Image…
★ 689
DFloat11
DFloat11 [NeurIPS '25]: Lossless Compression of LLMs and DiTs for Efficient GPU Inference
★ 642
wavemap
Fast, efficient and accurate multi-resolution, multi-sensor 3D occupancy mapping
★ 564
KVQuant
[NeurIPS 2024] KVQuant: Towards 10 Million Context Length LLM Inference with KV Cache Quantization
★ 431
Context-Engine
Context-Engine MCP - Agentic Context Compression Suite
★ 400
BK-SDM
A Compressed Stable Diffusion for Efficient Text-to-Image Generation [ECCV'24]
★ 319
picollm
On-device LLM Inference Powered by X-Bit Quantization
★ 314
distill
Context intelligence layer for LLM agents: persistent memory with write-time dedup, sensitivity tagging,…
★ 174
🔗 Famiglie affini

Misurato dai temi di GitHub condivisi da entrambi i progetti, ponderato in base a quanto è raro ciascun tema.