foundation-models

43 projects share this GitHub topic

foundation-models — LLaVA ★24.9kfoundation-modelsJanus — ★17.8kTabPFN — ★7.6kYuE — ★6.3kchronos-forecasting — ★5.6kLimiX — ★3.8kNExT-GPT — ★3.6kautodistill — ★2.7kInternVideo — ★2.3kMambaVision — ★2.2kalpaca_eval — ★2klag-llama — ★1.6kawesome-japanese-llm — ★1.4kKnowledgeEditingPapers — ★1.2kAwesome-Foundation-Models — ★1.2kFoundation-Models-Framework-Lab — ★1.2kONE-PEACE — ★1.1kPointLLM — ★1kVoxPoser — ★823ModelsGenesis — ★787GigaAM — ★688Awesome-Reasoning-Foundation-Models — ★656EmerNeRF — ★639Aether — ★603tokenize-anything — ★601Groma — ★585HoloMotion — ★583LLM4TS — ★567HPT — ★541TPA — ★459VLABench — ★450Olympus — ★428ai-powered-search — ★399BioReason — ★399MindVideo — ★389HyperSIGMA — ★379lmms-finetune — ★373CarDreamer — ★359fondant — ★358GRID-playground — ★345ViP-LLaVA — ★338Janus★ 17.8kTabPFN★ 7.6kYuE★ 6.3kchronos-forecasting★ 5.6kLimiX★ 3.8kNExT-GPT★ 3.6kautodistill★ 2.7kInternVideo★ 2.3kMambaVision★ 2.2kalpaca_eval★ 2klag-llama★ 1.6kawesome-japanese-llm★ 1.4kKnowledgeEditingPapers★ 1.2kAwesome-Foundation-Model…★ 1.2kFoundation-Models-Framew…★ 1.2kONE-PEACE★ 1.1kPointLLM★ 1kVoxPoser★ 823ModelsGenesis★ 787GigaAM★ 688Awesome-Reasoning-Founda…★ 656EmerNeRF★ 639Aether★ 603tokenize-anything★ 601Groma★ 585HoloMotion★ 583LLM4TS★ 567HPT★ 541TPA★ 459VLABench★ 450Olympus★ 428ai-powered-search★ 399BioReason★ 399MindVideo★ 389HyperSIGMA★ 379lmms-finetune★ 373CarDreamer★ 359fondant★ 358GRID-playground★ 345ViP-LLaVA★ 338

Lines connect members that are measurably related to each other. Dot size reflects stars.

🧬 Members
LLaVA
[NeurIPS'23 Oral] Visual Instruction Tuning (LLaVA) built towards GPT-4V level capabilities and beyond.
★ 24.9k
Janus
Janus-Series: Unified Multimodal Understanding and Generation Models
★ 17.8k
TabPFN
⚡ TabPFN: Foundation Model for Tabular Data ⚡
★ 7.6k
YuE
YuE: Open Full-song Music Generation Foundation Model, something similar to Suno.ai but open
★ 6.3k
chronos-forecasting
Chronos: Pretrained Models for Time Series Forecasting
★ 5.6k
LimiX
LimiX: Unleashing Structured-Data Modeling Capability for Generalist Intelligence…
★ 3.8k
NExT-GPT
Code and models for ICML 2024 paper, NExT-GPT: Any-to-Any Multimodal Large Language Model
★ 3.6k
autodistill
Images to inference with no labeling (use foundation models to train supervised models).
★ 2.7k
InternVideo
[ECCV2024] Video Foundation Models & Data for Multimodal Understanding
★ 2.3k
MambaVision
[CVPR 2025] Official PyTorch Implementation of MambaVision: A Hybrid Mamba-Transformer Vision Backbone
★ 2.2k
alpaca_eval
An automatic evaluator for instruction-following language models. Human-validated, high-quality, cheap, and…
★ 2k
lag-llama
Lag-Llama: Towards Foundation Models for Probabilistic Time Series Forecasting
★ 1.6k
awesome-japanese-llm
日本語LLMまとめ - Overview of Japanese LLMs
★ 1.4k
KnowledgeEditingPapers
Must-read Papers on Knowledge Editing for Large Language Models.
★ 1.2k
Awesome-Foundation-Models
A curated list of foundation models for vision and language tasks
★ 1.2k
Foundation-Models-Framework-Lab
A practical lab for building, testing, and evaluating apps with Apple's Foundation Models framework.
★ 1.2k
ONE-PEACE
A general representation model across vision, audio, language modalities. Paper: ONE-PEACE: Exploring One…
★ 1.1k
PointLLM
[ECCV 2024 Best Paper Candidate & TPAMI 2025] PointLLM: Empowering Large Language Models to Understand Point…
★ 1k
VoxPoser
VoxPoser: Composable 3D Value Maps for Robotic Manipulation with Language Models
★ 823
ModelsGenesis
[MICCAI 2019 Young Scientist Award] [MEDIA 2020 Best Paper Award] Models Genesis, one of the first…
★ 787
GigaAM
Foundational Model for Speech Recognition Tasks
★ 688
Awesome-Reasoning-Foundation-Models
✨✨Latest Papers and Benchmarks in Reasoning with Foundation Models
★ 656
EmerNeRF
PyTorch Implementation of EmerNeRF: Emergent Spatial-Temporal Scene Decomposition via Self-Supervision
★ 639
Aether
[ICCV 2025 & ICCV 2025 RIWM Outstanding Paper] Aether: Geometric-Aware Unified World Modeling
★ 603
tokenize-anything
[ECCV 2024] Tokenize Anything via Prompting
★ 601
Groma
[ECCV2024] Grounded Multimodal Large Language Model with Localized Visual Tokenization
★ 585
HoloMotion
HoloMotion: A Foundation Model for Whole-Body Humanoid Control
★ 583
LLM4TS
Large Language & Foundation Models for Time Series.
★ 567
HPT
Heterogeneous Pre-trained Transformer (HPT) as Scalable Policy Learner.
★ 541
TPA
[NeurIPS 2025 Spotlight] TPA: Tensor ProducT ATTenTion Transformer (https://arxiv.org/abs/2501.06425)
★ 459
VLABench
Official repo of VLABench, a large scale benchmark designed for fairly evaluating VLA, Embodied Agent, and…
★ 450
Olympus
[CVPR 2025 Highlight] Official code for "Olympus: A Universal Task Router for Computer Vision Tasks"
★ 428
ai-powered-search
The codebase for the book "AI-Powered Search" (Manning Publications, 2025) and associated "AI-Powered Search:…
★ 399
BioReason
BioReason: Incentivizing Multimodal Biological Reasoning within a DNA-LLM Model | NeurIPS '25
★ 399
MindVideo
Official code base for MinD-Video
★ 389
HyperSIGMA
The official repo for [TPAMI'25] "HyperSIGMA: Hyperspectral Intelligence Comprehension Foundation Model"
★ 379
lmms-finetune
A minimal codebase for finetuning large multimodal models, supporting llava-1.5/1.6, llava-interleave,…
★ 373
CarDreamer
World Model based Autonomous Driving Platform in CARLA :car:
★ 359
fondant
Production-ready data processing made easy and shareable
★ 358
GRID-playground
Platform for General Robot Intelligence Development
★ 345
ViP-LLaVA
[CVPR2024] ViP-LLaVA: Making Large Multimodal Models Understand Arbitrary Visual Prompts
★ 338
Awesome-Multimodal-LLM-Autonomous-Driving
[WACV 2024 Survey Paper] Multimodal Large Language Models for Autonomous Driving
★ 312
meta-prompting
Official implementation of Meta Prompting for AI Systems (https://arxiv.org/abs/2311.11482)
★ 307
🔗 Related families

Measured from GitHub topics shared by both projects, weighted by how rare each topic is.