attention

36 proyectos comparten este topic de GitHub

attention — annotated_deep_learning_paper_implementations ★67.1kattentionsglang — ★30.6knumpy-ml — ★16.3kOpenMythos — ★14.7kattention-is-all-you-need-pytorch — ★9.8kflashinfer — ★6kscenic — ★3.8kSageAttention — ★3.5kpytorch-GAT — ★2.7kawesome-multi-agent-papers — ★1.6klambda-networks — ★1.5ktransfusion-pytorch — ★1.4kgansformer — ★1.3kSimpleCVReproduction — ★1.3kself-attention-cv — ★1.2kperformer-pytorch — ★1.2kpytorch-original-transformer — ★1.1kgraphtransformer — ★1kNLP-Tutorials — ★950tiny-vllm — ★938rl4co — ★892VAD — ★869Medical-Transformer — ★861Deepdive-llama3-from-scratch — ★631keras_cv_attention_models — ★627neural_sp — ★594BiFormer — ★581ParaAttention — ★427Multimodal-Sentiment-Analysis — ★369CM3Leon — ★365seq2seq-summarizer — ★358probing-vits — ★341Chinese-ChatBot — ★324nanodl — ★306llm-flashcards — ★96TeaLeaves — ★42sglang★ 30.6knumpy-ml★ 16.3kOpenMythos★ 14.7kattention-is-all-you-nee…★ 9.8kflashinfer★ 6kscenic★ 3.8kSageAttention★ 3.5kpytorch-GAT★ 2.7kawesome-multi-agent-pape…★ 1.6klambda-networks★ 1.5ktransfusion-pytorch★ 1.4kgansformer★ 1.3kSimpleCVReproduction★ 1.3kself-attention-cv★ 1.2kperformer-pytorch★ 1.2kpytorch-original-transfo…★ 1.1kgraphtransformer★ 1kNLP-Tutorials★ 950tiny-vllm★ 938rl4co★ 892VAD★ 869Medical-Transformer★ 861Deepdive-llama3-from-scr…★ 631keras_cv_attention_model…★ 627neural_sp★ 594BiFormer★ 581ParaAttention★ 427Multimodal-Sentiment-Ana…★ 369CM3Leon★ 365seq2seq-summarizer★ 358probing-vits★ 341Chinese-ChatBot★ 324nanodl★ 306llm-flashcards★ 96TeaLeaves★ 42 · GitHub ↗

Las líneas conectan a los miembros que están mediblemente relacionados entre sí. El tamaño de los puntos refleja las estrellas.

🧬 Miembros
annotated_deep_learning_paper_implementations
🧑‍🏫 60+ Implementations/tutorials of deep learning papers with side-by-side notes 📝; including…
★ 67.1k
sglang
SGLang is a high-performance serving framework for large language models and multimodal models.
★ 30.6k
numpy-ml
Machine learning, in numpy
★ 16.3k
OpenMythos
A theoretical reconstruction of the Claude Mythos architecture, built from first principles using the…
★ 14.7k
attention-is-all-you-need-pytorch
A PyTorch implementation of the Transformer model in "Attention is All You Need".
★ 9.8k
flashinfer
FlashInfer: Kernel Library for LLM Serving
★ 6k
scenic
Scenic: A Jax Library for Computer Vision Research and Beyond
★ 3.8k
SageAttention
[ICLR2025, ICML2025, NeurIPS2025 Spotlight] Quantized Attention achieves speedup of 2-5x compared to…
★ 3.5k
pytorch-GAT
My implementation of the original GAT paper (Veličković et al.). I've additionally included the…
★ 2.7k
awesome-multi-agent-papers
A compilation of the best multi-agent papers
★ 1.6k
lambda-networks
Implementation of LambdaNetworks, a new approach to image recognition that reaches SOTA with less compute
★ 1.5k
transfusion-pytorch
Pytorch implementation of Transfusion, "Predict the Next Token and Diffuse Images with One Multi-Modal…
★ 1.4k
gansformer
Generative Adversarial Transformers
★ 1.3k
SimpleCVReproduction
Replication of simple CV Projects including attention, classification, detection, keypoint detection, etc.
★ 1.3k
self-attention-cv
Implementation of various self-attention mechanisms focused on computer vision. Ongoing repository.
★ 1.2k
performer-pytorch
An implementation of Performer, a linear attention-based transformer, in Pytorch
★ 1.2k
pytorch-original-transformer
My implementation of the original transformer model (Vaswani et al.). I've additionally included the…
★ 1.1k
graphtransformer
Graph Transformer Architecture. Source code for "A Generalization of Transformer Networks to Graphs",…
★ 1k
NLP-Tutorials
Simple implementations of NLP models. Tutorials are written in Chinese on my website https://mofanpy.com
★ 950
tiny-vllm
Build your own high performance LLM inference engine in C++ and CUDA - a smaller version of vLLM
★ 938
rl4co
A PyTorch library for all things Reinforcement Learning (RL) for Combinatorial Optimization (CO)
★ 892
VAD
Voice activity detection (VAD) toolkit including DNN, bDNN, LSTM and ACAM based VAD. We also provide our…
★ 869
Medical-Transformer
Official Pytorch Code for "Medical Transformer: Gated Axial-Attention for Medical Image Segmentation" -…
★ 861
Deepdive-llama3-from-scratch
Achieve the llama3 inference step-by-step, grasp the core concepts, master the process derivation, implement…
★ 631
keras_cv_attention_models
Keras beit,caformer,CMT,CoAtNet,convnext,davit,dino,efficientdet,edgenext,efficientformer,efficientnet,eva,fas…
★ 627
neural_sp
End-to-end ASR/LM implementation with PyTorch
★ 594
BiFormer
[CVPR 2023] Official code release of our paper "BiFormer: Vision Transformer with Bi-Level Routing Attention"
★ 581
ParaAttention
https://wavespeed.ai/ Context parallel attention that accelerates DiT model inference with dynamic caching
★ 427
Multimodal-Sentiment-Analysis
多模态情感分析——基于BERT+ResNet的多种融合方法
★ 369
CM3Leon
An open source implementation of "Scaling Autoregressive Multi-Modal Models: Pretraining and Instruction…
★ 365
seq2seq-summarizer
Pointer-generator reinforced seq2seq summarization in PyTorch
★ 358
probing-vits
Probing the representations of Vision Transformers.
★ 341
Chinese-ChatBot
中文聊天机器人,基于10万组对白训练而成,采用注意力机制,对一般问题都会生成…
★ 324
nanodl
JAX library for training sub-4B foundation models for edge
★ 306
llm-flashcards
Visual knowledge bank for understanding large language models, with 180 concept cards from tokenization to…
★ 96
TeaLeaves
End-to-end pipeline for seeing how LLMs actually process your prompts. Capture attention across every layer,…
★ 42 · GitHub ↗
🔗 Familias relacionadas

Medido a partir de los temas de GitHub compartidos por ambos proyectos, ponderado por cuán raros son cada uno de los temas.