Skip to content
repository radar
Models & Inference

🧠Oumi

ACTIVE STEADY

oumi-ai/oumi · homepage ↗

Easily fine-tune, evaluate and deploy Qwen, Gemma, or any open weight LLM!

⭐ Popular: 9k stars, gaining about 6 a week

View on GitHub ↗

repo profile

vintage 2024 2 years old
delivery product
language Python
license Apache-2.0

momentum

total stars 9k
stars added last week +6
weekly downloads 477 /wk · PyPI
commits / week 5 steady · -5% vs prior mo
issues closed 100% ~14d to close, median
contributors 52
release cadence monthly
last activity yesterday

durability

backing VC-backed Oumi
openness permissive
bus factor 3 concentrated
top-author share 27% 6 mo

bus factor = how many people it takes to cover more than half the commits (6 months). 1 is a solo project; higher means the work is spread across a team. top-author share is the single busiest author's slice of those commits.

since we covered it

monthly average + 120/mo (+2%/mo) · + 2k total since PR#2

why it's a big deal

  • Enables training and fine-tuning of large-scale AI models with support for techniques like LoRA, QLoRA, and DPO.
  • Works across multiple model architectures, including Llama, DeepSeek, Qwen, and Phi.
  • Integrates seamlessly with cloud providers (AWS, Azure, GCP, Lambda) for remote job execution.

under the hood

  • Supports zero-boilerplate configuration for fine-tuning, distillation, and benchmarking.
  • Includes native tools for LLM-as-a-judge, data synthesis, and structured evaluation.
  • Runs efficiently on GPUs and NPUs, leveraging distributed training techniques.Designed for both research and enterprise AI model development.

our take from PR#2, 2025-02-19

star history

PR#2 · 7k9k now May 2024Aug 2026
  1. PR#2 7k 2025-02-19
  2. now 9k + 2k since first covered

curve is sampled from GitHub's star history, plus our own daily readings since we covered it; the dashed stretch is before we first covered it, the solid line since. figures at coverage are the numbers we printed then (approx.), current count is live.

understory

Quietly building: more output than attention, for now.

+20 understory score output 62 · clout 42
Aug 2025 Jul 2026
  • output, commits & releases
  • clout, star velocity

output = commits & releases; clout = star velocity, both 0 to 100 monthly indices; the gap where output runs above clout is the understory. The understory →

covered in

  • PR#2 2025-02-19 on the radar

    The End-to-End Platform for Training AI Foundation Models

similar projects

compare these →
  • 🦙 LLaMA-Factory

    7.9× the stars

    Unified Efficient Fine-Tuning of 100+ LLMs & VLMs (ACL 2024)

    74k ACTIVE
  • Unsloth

    7.6× the stars

    Local UI to run and train LLMs and diffusion models, including Qwen3.8, Kimi K3, MiniMax-H3, Gemma 4, DeepSeek-V4, FLUX and more.

    72k ACTIVE
  • 🧩 TensorZero

    Rust

    TensorZero is an open-source LLMOps platform that unifies an LLM gateway, observability, evaluation, optimization, and experimentation.

    12k OFF THE RADAR

comments

Sign in with GitHub to add your blip on Oumi.

loading comments…