Models, Training & Serving

65 starred repositories in this area.

65 result(s)

rasbt/LLMs-from-scratch

Implement a ChatGPT-like LLM in PyTorch from scratch, step by step

★ 104.6k · Jupyter Notebook · updated 2026-09

Jupyter Notebookaiartificial-intelligenceattention-mechanism

deepseek-ai/DeepSeek-R1

DeepSeek 的開源推理模型,用大規模強化學習訓練出接近 o1 等級的思考能力。

★ 92k · — · updated 2025-06

unslothai/unsloth

Local UI to run and train LLMs and diffusion models. Supports GGUF, MLX, Qwen3.8, DeepSeek-V4, MiniMax-H3, Gemma 4, FLUX and more.

★ 75.8k · Python · updated 2026-09

Pythonagentaichatgpt

hiyouga/LlamaFactory

Unified Efficient Fine-Tuning of 100+ LLMs & VLMs (ACL 2024)

★ 74.6k · Python · updated 2026-09

Pythonagentaideepseek

jingyaogong/minimind

🧠 Train a 64M-parameter LLM from scratch in just 2h!

★ 59.8k · Python · updated 2026-09

Pythonartificial-intelligencelarge-language-model

google/langextract

A Python library for extracting structured information from unstructured text using LLMs with precise source grounding and interactive visualization.

★ 38.5k · Python · updated 2026-09

Pythongeminigemini-aigemini-api

stanfordnlp/dspy

DSPy: The framework for programming—not prompting—language models

★ 37.9k · Python · updated 2026-09

Python

alchaincyf/nuwa-skill

你想蒸馏的下一个员工,何必是同事。蒸馏任何人的思维方式——心智模型、决策启发式、表达DNA。Distill how anyone thinks.

★ 32.3k · Python · updated 2026-08

Python

QwenLM/Qwen3

Qwen3 is the large language model series developed by Qwen team, Alibaba Cloud.

★ 27.6k · Python · updated 2026-01

Python

huggingface/open-r1

Fully open reproduction of DeepSeek-R1

★ 26.5k · Python · updated 2026-04

Python

confident-ai/deepeval

The LLM Evaluation Framework

★ 18.2k · Python · updated 2026-09

Pythonevaluation-frameworkevaluation-metricsllm-evaluation

dair-ai/ML-YouTube-Courses

📺 Discover the latest machine learning / AI courses on YouTube.

★ 17.4k · — · updated 2024-01

aidata-sciencedeep-learning

ageron/handson-ml3

A series of Jupyter notebooks that walk you through the fundamentals of Machine Learning and Deep Learning in Python using Scikit-Learn, Keras and TensorFlow 2.

★ 14.1k · Jupyter Notebook · updated 2026-05

Jupyter Notebook

Lightning-AI/litgpt

20+ high-performance LLMs with recipes to pretrain, finetune and deploy at scale.

★ 13.7k · Python · updated 2026-09

Pythonaiartificial-intelligencedeep-learning

dair-ai/AI-Papers-of-the-Week

🔥Highlighting the top ML papers every week.

★ 13.2k · — · updated 2026-09

aidata-sciencedeeplearning

ShishirPatil/gorilla

Gorilla: Training and Evaluating LLMs for Function Calls (Tool Calls)

★ 13k · Python · updated 2026-04

Pythonapiapi-documentationchatgpt

NVIDIA/personaplex

PersonaPlex code.

★ 10.5k · Python · updated 2026-03

Python

xorbitsai/inference

Swap GPT for any LLM by changing a single line of code. Xinference lets you run open-source, speech, and multimodal models on cloud, on-prem, or your laptop — all through one unified, production-ready inference API.

★ 9.6k · Python · updated 2026-09

Pythonartificial-intelligencedeploymentdiffusers

oumi-ai/oumi

Easily fine-tune, evaluate and deploy Qwen, Gemma, or any open weight LLM!

★ 9.4k · Python · updated 2026-09

Pythondpoevaluationfine-tuning

NVIDIA/garak

the LLM vulnerability scanner

★ 9.1k · Python · updated 2026-09

Pythonaillm-evaluationllm-security

NVIDIA/Isaac-GR00T

NVIDIA Isaac GR00T N1.7 - A Foundation Model for Generalist Robots.

★ 8k · Python · updated 2026-08

Python

ai-dynamo/dynamo

A Datacenter Scale Distributed Inference Serving Framework

★ 8k · Rust · updated 2026-09

Rustdiffusiondisaggregated-servingkubernetes

microsoft/TinyTroupe

LLM-powered multiagent persona simulation for imagination enhancement and business insights.

★ 7.6k · Jupyter Notebook · updated 2026-07

Jupyter Notebook

guardrails-ai/guardrails

Adding guardrails to large language models.

★ 7.4k · Python · updated 2026-09

Pythonaifoundation-modelgpt-3

linkedin/Liger-Kernel

Efficient Triton Kernels for LLM Training

★ 6.6k · Python · updated 2026-09

Pythonfinetuninggemma2hacktoberfest

jeinlee1991/chinese-llm-benchmark

非线智能 NoneLinear - ReLE评测:中文AI大模型能力评测(持续更新):目前已囊括374个大模型,覆盖chatgpt、gpt-5.4、谷歌gemini-3.1-pro、Claude-4.6、文心ERNIE-X1.1、ERNIE-5.0、qwen3.6-max、qwen3.6-plus、百川、讯飞星火、商汤senseChat等商用模型, 以及step3.5-flash、kimi-k2.6、ernie4.5、MiniMax-M2.7、deepseek-v4、Qwen3.6、llama4、智谱GLM-5.1、MiMo-V2、LongCat、gemma4、mistral等开源大模型。不仅提供排行榜,也提供规模超200万的大模型缺陷库!方便广大社区研究分析、改进大模型。

★ 6.4k · — · updated 2026-09

agentic-aiartificial-intelligencellm-agent

shibing624/MedicalGPT

MedicalGPT: Training Your Own Medical GPT Model with ChatGPT Training Pipeline. 训练医疗大模型,实现了包括增量预训练(PT)、有监督微调(SFT)、RLHF、DPO、ORPO、GRPO。

★ 5.8k · Python · updated 2026-06

Pythonchatgptdpogpt

getzep/zep

Zep | Examples, Integrations, & More

★ 4.9k · Python · updated 2026-09

Pythonaiknowledge-graphslanguage-model

openai/simple-evals

OpenAI 官方的輕量模型評測腳本,用來重現各家模型的標準 benchmark 分數。

★ 4.6k · Python · updated 2026-04

Python

PrimeIntellect-ai/verifiers

Our library for RL environments + evals

★ 4.6k · Python · updated 2026-09

Python

build-with-groq/g1

g1: Using Llama-3.1 70b on Groq to create o1-like reasoning chains

★ 4.2k · Python · updated 2025-12

Pythonmanaged-by-terraform

predibase/lorax

Multi-LoRA inference server that scales to 1000s of fine-tuned LLMs

★ 3.8k · Python · updated 2026-05

Pythonfine-tuninggptllama

ridgerchu/matmulfreellm

Implementation for MatMul-free LM.

★ 3.1k · Python · updated 2026-09

Pythonlarge-language-modellinear-transformerllm

zjunlp/EasyEdit

[ACL 2024] An Easy-to-use Knowledge Editing Framework for LLMs.

★ 2.9k · Jupyter Notebook · updated 2026-07

Jupyter Notebookartificial-intelligencebaichuanchatgpt

WenjieDu/PyPOTS

A Python toolkit/library for reality-centric machine/deep learning & data mining on partially-observed time series, with 50+ SOTA neural network models for scientific analysis tasks (imputation, classification, clustering, forecasting, anomaly detection, cleaning) on incomplete industrial irregularly-sampled multivariate TS with NaN missing values

★ 2.1k · Python · updated 2026-09

Pythonanomaly-detectionclassificationclustering

meta-llama/synthetic-data-kit

Tool for generating high quality Synthetic datasets

★ 1.6k · Python · updated 2025-10

Pythondatagenerationllm

NVIDIA/RULER

This repo contains the source code for RULER: What’s the Real Context Size of Your Long-Context Language Models?

★ 1.6k · Python · updated 2026-07

Python

AlibabaResearch/DAMO-ConvAI

DAMO-ConvAI: The official repository which contains the codebase for Alibaba DAMO Conversational AI.

★ 1.6k · Python · updated 2026-06

Pythonconversational-aideep-learningdialog

Google-Health/medgemma

Google Health 的 MedGemma 醫療多模態開放模型,能讀醫學影像與病歷文字。

★ 1.6k · Jupyter Notebook · updated 2026-06

Jupyter Notebook

sail-sg/understand-r1-zero

Understanding R1-Zero-Like Training: A Critical Perspective

★ 1.3k · Python · updated 2025-08

Pythonllmr1-zeroreasoning

SakanaAI/self-adaptive-llms

A Self-adaptation Framework🐙 that adapts LLMs for unseen tasks in real-time!

★ 1.2k · Python · updated 2025-01

Python

cxcscmu/Craw4LLM

Official repository for "Craw4LLM: Efficient Web Crawling for LLM Pretraining"

★ 664 · Python · updated 2025-02

Pythoncrawlercrawlinglarge-language-models

Wang-ML-Lab/llm-continual-learning-survey

[CSUR 2025] Continual Learning of Large Language Models: A Comprehensive Survey

★ 562 · — · updated 2025-12

continual-learninglarge-language-modelllm

ZongqianLi/ReasonGraph

[ACL 2025 Demo] Repository for the demo and paper: ReasonGraph: Visualisation of Reasoning Paths

★ 516 · HTML · updated 2026-03

HTML

NVIDIA/Star-Attention

Efficient LLM Inference over Long Sequences

★ 392 · Python · updated 2025-06

Pythonattention-mechanismlarge-language-modelsllm-inference

SamuelSchmidgall/AgentClinic

Agent benchmark for medical diagnosis

★ 352 · Python · updated 2024-12

Python

shangshang-wang/Tina

[ICLR 2026] Tina: Tiny Reasoning Models via LoRA

★ 337 · Python · updated 2025-09

Python

mtkresearch/MR-Models

聯發創新基地(MediaTek Research) 致力於研究基礎模型。我們將研究體現在適合繁體中文使用者的模型上,並在使用權許可的情況下,提供模型給學術界研究或產業界使用。

★ 295 · — · updated 2026-04

llm

knowsuchagency/promptic

90% of what you need for LLM app development. Nothing you don't.

★ 274 · Python · updated 2025-08

Pythonaillmspython

Jiaqi-Chen-00/ImBD

[AAAI 2025 oral] Official repository of Imitate Before Detect: Aligning Machine Stylistic Preference for Machine-Revised Text Detection

★ 264 · Python · updated 2025-04

Pythonai-content-detectorai-safetyllm-detection

patrickloeber/llm-data-scrapers

A list of useful Open Source tools and scrapers to gather data for LLMs

★ 248 · — · updated 2025-02

CryptoAILab/JailbreakEval

[NDSS'25 Best Technical Poster] A collection of automated evaluators for assessing jailbreak attempts.

★ 196 · Python · updated 2025-04

Pythonllm-jailbreaksllm-safety

LibertFan/AI_Hospital

AI Hospital: Interactive Evaluation and Collaboration of LLMs as Intern Doctors for Clinical Diagnosis

★ 195 · Python · updated 2024-09

Python

FreedomIntelligence/Chain-of-Diagnosis

An interpretable large language model (LLM) for medical diagnosis.

★ 162 · Python · updated 2024-09

Python

ai-twinkle/Eval

High-performance LLM evaluation framework with parallel API calls — up to 17× faster than sequential tools. Supports box, math, and logit-based evaluation.

★ 109 · Python · updated 2026-08

Pythonevalevaluationllm

stream-bench/stream-bench

We propose a pioneering benchmark to evaluate LLM agents' ability to improve over time in streaming scenarios

★ 85 · Python · updated 2024-10

Python

LLMSELECTOR/LLMSELECTOR

替每個查詢挑選最合適模型的路由研究,兼顧成本與品質。

★ 81 · Jupyter Notebook · updated 2025-09

Jupyter Notebook

envy-ai/Wan2.1-quantized

Wan2.1, quantized and optimized so it fits on your 3090/4090

★ 34 · Python · updated 2025-02

Python

LiuYuWei/llm-colab-application

[ 演講用 ] LLM AI 模型 Colab 應用

★ 27 · Jupyter Notebook · updated 2025-12

Jupyter Notebook

AIAnytime/Qwen2-VL-Fine-Tuning

Qwen2 VL Fine Tuning using Llama Factory

★ 19 · Jupyter Notebook · updated 2024-09

Jupyter Notebook

Jakevin/o1-like-prompt

Reasoning chain prompt

★ 19 · — · updated 2024-09

DanielSun94/conversational_diagnosis

對話式醫療診斷的研究程式碼。

★ 8 · Python · updated 2024-07

Python

appier-research/streambench-final-project

The final project for Applied Deep Learning (ADL) 2024 @NTU lectured by Prof. Yun-Nung (Vivian) Chen. The project is based on the paper StreamBench: Towards Benchmarking Continuous Improvement of Language Agents.

★ 7 · Python · updated 2024-12

Python

TGoldsack1/BioLaySumm2024-evaluation_scripts

Evaluation script for the BioLaySumm 2024 shared task.

★ 6 · Python · updated 2024-02

Python

nullxjx/transformers-online-inference

使用 transformers 和 fastapi 搭建的LLM在线推理服务

★ 4 · Python · updated 2024-09

Python

中文版