gpu

51 Projekte teilen dieses GitHub-Topic

gpu — pytorch ★101.7kgpufastai — ★28.1kWebGL-Fluid-Simulation — ★16.5ktvm — ★13.6kflashinfer — ★6kcuml — ★5.2klemonade — ★5kkoharu — ★4.9kMegEngine — ★4.8kexecutorch — ★4.8ktiny-cuda-nn — ★4.5kdeepflow — ★4.2kllm-d — ★3.8kjittor — ★3.2kchitu — ★3.1kheavydb — ★3.1kleptonai — ★2.8kdifftaichi — ★2.7ktorchrec — ★2.6kdeepdetect — ★2.6kopenlake — ★2.2ktrainer — ★2.2kglim — ★1.7kgpu-io — ★1.5kisaac_ros_visual_slam — ★1.4kgpu_poor — ★1.4kmodal-examples — ★1.2kwebgl-wind — ★1.1kTalos — ★990femtoGPT — ★935GPUMD — ★811can-i-finetune-this — ★790isaac_ros_nvblox — ★724DFloat11 — ★642VerletIntegration — ★581neurokernel — ★565popsift — ★498blub — ★483isaac_ros_pose_estimation — ★475warpx — ★470zinc — ★467fastai★ 28.1kWebGL-Fluid-Simulation★ 16.5ktvm★ 13.6kflashinfer★ 6kcuml★ 5.2klemonade★ 5kkoharu★ 4.9kMegEngine★ 4.8kexecutorch★ 4.8ktiny-cuda-nn★ 4.5kdeepflow★ 4.2kllm-d★ 3.8kjittor★ 3.2kchitu★ 3.1kheavydb★ 3.1kleptonai★ 2.8kdifftaichi★ 2.7ktorchrec★ 2.6kdeepdetect★ 2.6kopenlake★ 2.2ktrainer★ 2.2kglim★ 1.7kgpu-io★ 1.5kisaac_ros_visual_slam★ 1.4kgpu_poor★ 1.4kmodal-examples★ 1.2kwebgl-wind★ 1.1kTalos★ 990femtoGPT★ 935GPUMD★ 811can-i-finetune-this★ 790isaac_ros_nvblox★ 724DFloat11★ 642VerletIntegration★ 581neurokernel★ 565popsift★ 498blub★ 483isaac_ros_pose_estimatio…★ 475warpx★ 470zinc★ 467

Linien verbinden Mitglieder, die messbar miteinander verwandt sind. Die Punktgröße spiegelt die Sterne wider.

🧬 Mitglieder
pytorch
Tensors and Dynamic neural networks in Python with strong GPU acceleration
★ 101.7k
fastai
The fastai deep learning library
★ 28.1k
WebGL-Fluid-Simulation
Play with fluids in your browser (works even on mobile)
★ 16.5k
tvm
Open Machine Learning Compiler Framework
★ 13.6k
flashinfer
FlashInfer: Kernel Library for LLM Serving
★ 6k
cuml
cuML - RAPIDS Machine Learning Library
★ 5.2k
lemonade
Lemonade helps users discover and run local AI apps by serving optimized LLMs right from their own GPUs and…
★ 5k
koharu
ML-powered manga translator, written in Rust.
★ 4.9k
MegEngine
MegEngine 是一个快速、可拓展、易于使用且支持自动求导的深度学习框架
★ 4.8k
executorch
On-device AI across mobile, embedded and edge for PyTorch
★ 4.8k
tiny-cuda-nn
Lightning fast C++/CUDA neural network framework
★ 4.5k
deepflow
eBPF Observability - Distributed Tracing and Profiling
★ 4.2k
llm-d
Achieve state of the art inference performance with modern accelerators on Kubernetes
★ 3.8k
jittor
Jittor is a high-performance deep learning framework based on JIT compiling and meta-operators.
★ 3.2k
chitu
High-performance inference framework for large language models, focusing on efficiency, flexibility, and…
★ 3.1k
heavydb
HeavyDB (formerly MapD/OmniSciDB)
★ 3.1k
leptonai
A Pythonic framework to simplify AI service building
★ 2.8k
difftaichi
10 differentiable physical simulators built with Taichi differentiable programming (DiffTaichi, ICLR 2020)
★ 2.7k
torchrec
Pytorch domain library for recommendation systems
★ 2.6k
deepdetect
Deep Learning Server and CLI for Torch and TensorRT
★ 2.6k
openlake
OpenLake is a high performance storage engine for efficient LLM inference and GPU Training
★ 2.2k
trainer
Distributed AI Model Training and LLM Fine-Tuning on Kubernetes
★ 2.2k
glim
GLIM: versatile and extensible point cloud-based 3D localization and mapping framework
★ 1.7k
gpu-io
A GPU-accelerated computing library for running physics simulations and other GPGPU computations in a web…
★ 1.5k
isaac_ros_visual_slam
Visual SLAM/odometry package based on NVIDIA-accelerated cuVSLAM
★ 1.4k
gpu_poor
Calculate token/s & GPU memory requirement for any LLM. Supports llama.cpp/ggml/bnb/QLoRA quantization
★ 1.4k
modal-examples
Examples of programs built using Modal
★ 1.2k
webgl-wind
Wind power visualization with WebGL particles
★ 1.1k
Talos
GPU worker client for the Talos network. Pairs with your Talos account, serves open-model inference jobs over…
★ 990
femtoGPT
Pure Rust implementation of a minimal Generative Pretrained Transformer
★ 935
GPUMD
Graphics Processing Units Molecular Dynamics
★ 811
can-i-finetune-this
Estimate whether a Hugging Face model fits and fine-tunes on your local GPU.
★ 790
isaac_ros_nvblox
NVIDIA-accelerated 3D scene reconstruction and Nav2 local costmap provider using nvblox
★ 724
DFloat11
DFloat11 [NeurIPS '25]: Lossless Compression of LLMs and DiTs for Efficient GPU Inference
★ 642
VerletIntegration
A real-time particle simulation that uses Verlet Integration
★ 581
neurokernel
Neurokernel Project
★ 565
popsift
PopSift is an implementation of the SIFT algorithm in CUDA.
★ 498
blub
3D fluid simulation experiments in Rust, using WebGPU-rs (WIP)
★ 483
isaac_ros_pose_estimation
Deep learned, NVIDIA-accelerated 3D object pose estimation
★ 475
warpx
WarpX is an advanced Particle-In-Cell code.
★ 470
zinc
Zig INferenCe Engine — Local LLM inference on AMD GPUs and Apple Silicon
★ 467
cucim
cuCIM - RAPIDS GPU-accelerated image processing library
★ 463
JetStream
JetStream is a throughput and memory optimized engine for LLM inference on XLA devices, starting with TPUs…
★ 451
hoomd-blue
Molecular dynamics and Monte Carlo soft matter simulation on GPUs.
★ 444
rag-chatbot
RAG (Retrieval-augmented generation) ChatBot that provides answers based on contextual information extracted…
★ 433
WaterBall
Fluid simulation on a sphere🌏
★ 390
knowhere
Vector search engine inside Milvus, integrating FAISS, HNSW, DiskANN.
★ 371
SPH_Taichi
A high-performance implementation of SPH in Taichi.
★ 319
mppi_numba
A GPU implementation of Model Predictive Path Integral (MPPI) control that uses a probabilistic…
★ 310
isaac_ros_common
Common utilities, packages, scripts, and testing infrastructure for Isaac ROS packages.
★ 307
attyx
GPU accelerated terminal for agentic workflows
★ 228
🔗 Verwandte Familien

Gemessen anhand der von beiden Projekten geteilten GitHub-Themen, gewichtet nach der Seltenheit jedes Themas.