Home
Browse all repos
/
Discover and explore top open-source AI tools and projects—updated daily.
Home
Browse all repos
Home
>
Users
>
Woosuk Kwon
Woosuk Kwon
Coauthor of vLLM
GitHub
Starred Projects (68)
uccl
by
uccl-project
0.1%
2k
GPU collective communication library for ML workloads
Starred by
+1
Created 1 year ago
Updated 5 days ago
ml-engineering
by
stas00
0.1%
19k
Open book for LLM/VLM training engineers
Starred by
+19
Created 6 years ago
Updated 2 days ago
torchtitan
by
pytorch
0%
6k
PyTorch platform for generative AI model training research
Starred by
+12
Created 2 years ago
Updated 9 hours ago
vllm-omni
by
vllm-project
0.1%
7k
Omni-modality model inference and serving framework
Starred by
Created 1 year ago
Updated 8 hours ago
verl
by
verl-project
0.0%
24k
RL training library for LLMs
Starred by
+16
Created 1 year ago
Updated 7 hours ago
SkyRL
by
NovaSky-AI
0.1%
2k
RL training pipeline for multi-turn tool use LLMs, optimized for real-world tasks
Starred by
+16
Created 1 year ago
Updated 14 hours ago
tinker-cookbook
by
thinking-machines-lab
0.1%
4k
Advanced LLM fine-tuning SDK and example cookbook
Starred by
+8
Created 1 year ago
Updated 18 hours ago
batch_invariant_ops
by
thinking-machines-lab
0.1%
1k
Enhance LLM inference determinism
Starred by
+1
Created 1 year ago
Updated 11 months ago
recipes
by
vllm-project
0%
1k
LLM inference recipes
Starred by
Created 1 year ago
Updated 1 day ago
openevolve
by
algorithmicsuperintelligence
0.0%
8k
Coding agent for scientific/algorithmic discovery, based on AlphaEvolve paper
Starred by
+6
Created 1 year ago
Updated 8 hours ago
nano-vllm
by
GeeeekExplorer
0.0%
16k
Lightweight vLLM implementation from scratch
Starred by
+2
Created 1 year ago
Updated 5 months ago
dynamo
by
ai-dynamo
0.0%
8k
Inference framework for distributed generative AI model serving
Starred by
+7
Created 1 year ago
Updated 7 hours ago
ArcticInference
by
snowflakedb
0%
488
vLLM plugin for high-throughput, low-latency LLM and embedding inference
Starred by
Created 1 year ago
Updated 2 weeks ago
MiMo
by
XiaomiMiMo
0.0%
2k
LLM for reasoning, pre-trained and post-trained for math/code tasks
Starred by
Created 1 year ago
Updated 1 year ago
rllm
by
rllm-org
0%
6k
Framework for post-training language agents via reinforcement learning
Starred by
+2
Created 1 year ago
Updated 1 day ago
chatgpt_system_prompt
by
LouisShark
0%
11k
GPT system prompt collection for prompt engineering and security education
Starred by
+2
Created 2 years ago
Updated 2 days ago
fairseq2
by
facebookresearch
0%
1k
Sequence modeling toolkit for content generation research
Created 3 years ago
Updated 1 month ago
Mooncake
by
kvcache-ai
0.0%
7k
Research paper on a disaggregated architecture for LLM serving
Starred by
+2
Created 2 years ago
Updated 1 day ago
xgrammar
by
mlc-ai
0%
2k
Library for efficient structured generation
Starred by
+5
Created 2 years ago
Updated 1 day ago
Liger-Kernel
by
linkedin
0%
7k
Triton kernels for efficient LLM training
Starred by
+9
Created 2 years ago
Updated 1 day ago
Nanoflow
by
efeslab
0%
975
LLM serving framework for high throughput
Starred by
Created 2 years ago
Updated 6 months ago
xla
by
pytorch
0%
3k
PyTorch on XLA devices
Starred by
+15
Created 8 years ago
Updated 1 day ago
ao
by
pytorch
0%
3k
PyTorch library for quantization and sparsity in training/inference
Starred by
+11
Created 2 years ago
Updated 10 hours ago
llm-compressor
by
vllm-project
0.0%
4k
Transformers-compatible library for LLM compression, optimized for vLLM deployment
Starred by
+3
Created 2 years ago
Updated 21 hours ago
intel-extension-for-pytorch
by
intel
0%
2k
PyTorch extension for performance boost on Intel platforms
Starred by
Created 6 years ago
Updated 6 months ago
ThunderKittens
by
HazyResearch
0.0%
4k
CUDA kernel framework for fast deep learning primitives
Starred by
+16
Created 2 years ago
Updated 3 weeks ago
mirage
by
mirage-project
0.1%
3k
Tool for fast GPU kernel generation via superoptimization
Starred by
+1
Created 2 years ago
Updated 2 days ago
AutoAWQ
by
casper-hansen
0%
2k
AutoAWQ is a tool for 4-bit quantized LLM inference
Starred by
+5
Created 3 years ago
Updated 1 year ago
grok-1
by
xai-org
0%
52k
JAX example code for loading and running Grok-1 open-weights model
Starred by
+22
Created 2 years ago
Updated 2 years ago
lm-evaluation-harness
by
EleutherAI
0.0%
14k
Framework for few-shot language model evaluation
Starred by
+18
Created 6 years ago
Updated 3 weeks ago
LLMSys-PaperList
by
AmberLJC
0%
2k
Curated list of LLM systems papers
Starred by
Created 3 years ago
Updated 2 months ago
aici
by
microsoft
0%
2k
AICI constrains LLM output using (Wasm) programs
Starred by
+7
Created 3 years ago
Updated 1 year ago
mlc-llm
by
mlc-ai
0%
23k
Universal LLM deployment engine with ML compilation
Starred by
+21
Created 3 years ago
Updated 3 days ago
mscclpp
by
microsoft
0%
562
GPU-driven communication stack for scalable AI applications
Starred by
Created 3 years ago
Updated 15 hours ago
sglang
by
sgl-project
0.0%
37k
Fast serving framework for LLMs and vision language models
Starred by
+36
Created 2 years ago
Updated 9 hours ago
flashinfer
by
flashinfer-ai
0.1%
7k
Kernel library for LLM serving
Starred by
+13
Created 3 years ago
Updated 11 hours ago
punica
by
punica-ai
0%
1k
LoRA serving system (research paper) for multi-tenant LLM inference
Starred by
+3
Created 3 years ago
Updated 2 years ago
LLMCompiler
by
SqueezeAILab
0%
2k
LLM compiler for parallel function calling
Starred by
+2
Created 2 years ago
Updated 2 years ago
gpt-fast
by
meta-pytorch
0%
6k
PyTorch text generation for efficient transformer inference
Starred by
+20
Created 3 years ago
Updated 1 year ago
TensorRT-LLM
by
NVIDIA
0.0%
15k
LLM inference optimization SDK for NVIDIA GPUs
Starred by
+18
Created 3 years ago
Updated 10 hours ago
WizardLM
by
nlpxucan
0%
9k
LLMs built using Evol-Instruct for complex instruction following
Starred by
+15
Created 3 years ago
Updated 1 year ago
outlines
by
dottxt-ai
0%
16k
SDK for structured LLM text generation
Starred by
+34
Created 3 years ago
Updated 2 weeks ago
Awesome-LLM
by
Hannibal046
0%
27k
Curated list of Large Language Model resources
Starred by
+8
Created 3 years ago
Updated 1 year ago
gorilla
by
ShishirPatil
0.0%
13k
LLM tool-use framework for API invocation and function calling
Starred by
+15
Created 3 years ago
Updated 6 months ago
LLMSurvey
by
RUCAIBox
0%
12k
Survey paper for large language models
Starred by
+2
Created 3 years ago
Updated 1 year ago
CTranslate2
by
OpenNMT
0.0%
5k
Fast inference engine for Transformer models
Starred by
+6
Created 7 years ago
Updated 5 days ago
SqueezeLLM
by
SqueezeAILab
0%
724
Quantization framework for efficient LLM serving (ICML 2024 paper)
Starred by
Created 3 years ago
Updated 2 years ago
vllm
by
vllm-project
0.0%
93k
LLM serving engine for high-throughput, memory-efficient inference
Starred by
+59
Created 3 years ago
Updated 12 hours ago
Awesome-LLMOps
by
tensorchord
0%
6k
Curated list of LLMOps tools for developers
Starred by
+3
Created 4 years ago
Updated 5 days ago
FastChat
by
lm-sys
0%
40k
Open platform for training, serving, and evaluating LLM-based chatbots
Starred by
+36
Created 3 years ago
Updated 5 months ago
llama
by
meta-llama
0%
60k
Inference code for Llama 2 models (deprecated)
Starred by
+38
Created 3 years ago
Updated 1 year ago
Megatron-LM
by
NVIDIA
0.0%
18k
Framework for training transformer models at scale
Starred by
+21
Created 7 years ago
Updated 7 hours ago
flash-attention
by
Dao-AILab
0.0%
25k
Fast, memory-efficient attention implementation
Starred by
+31
Created 4 years ago
Updated 3 days ago
TransformerEngine
by
NVIDIA
0%
4k
Library for Transformer model acceleration on NVIDIA GPUs
Starred by
+5
Created 4 years ago
Updated 19 hours ago
AITemplate
by
facebookincubator
0%
5k
Generate high-performance inference engines
Starred by
+19
Created 4 years ago
Updated 3 days ago
x-transformers
by
lucidrains
0%
6k
Transformer library with extensive experimental features
Starred by
+7
Created 6 years ago
Updated 5 days ago
compiler-and-arch
by
KnowingNothing
0%
538
Compiler/architecture resources for emerging domains
Starred by
Created 4 years ago
Updated 1 year ago
skypilot
by
skypilot-org
0%
11k
Framework for cloud AI/batch jobs, unifying execution across diverse infrastructure
Starred by
+25
Created 5 years ago
Updated 18 hours ago
metaseq
by
facebookresearch
0%
7k
Codebase for large-scale transformer model development and deployment
Starred by
+11
Created 4 years ago
Updated 2 years ago
FasterTransformer
by
NVIDIA
0%
6k
Optimized transformer library for inference
Starred by
+12
Created 5 years ago
Updated 2 years ago
alpa
by
alpa-projects
0%
3k
Auto-parallelization framework for large-scale neural network training and serving
Starred by
+17
Created 5 years ago
Updated 2 years ago
transformers
by
huggingface
0.0%
167k
ML library for pretrained model inference and training
Starred by
+96
Created 8 years ago
Updated 22 hours ago
ray
by
ray-project
0%
44k
AI compute engine for scaling Python and AI applications
Starred by
+53
Created 10 years ago
Updated 17 hours ago
awesome-tensor-compilers
by
merrymercy
0%
3k
Curated list of tensor compiler projects and papers
Starred by
+10
Created 6 years ago
Updated 2 years ago
tvm
by
apache
0%
14k
Compiler stack for deep learning systems
Starred by
+20
Created 10 years ago
Updated 22 hours ago
cutlass
by
NVIDIA
0.0%
11k
CUDA C++ and Python DSLs for high-performance linear algebra
Starred by
+23
Created 8 years ago
Updated 2 weeks ago
TensorRT
by
NVIDIA
0.0%
13k
SDK for accelerated deep learning inference on NVIDIA GPUs
Starred by
+4
Created 7 years ago
Updated 2 weeks ago
DeepLearningExamples
by
NVIDIA
0%
15k
Deep learning examples for training and deployment
Starred by
+8
Created 8 years ago
Updated 2 years ago
Feedback? Help us improve.