AI Year in Review · Oct 10, 2025 – Oct 10, 2026

Open-source models & trending GitHub repos

182 stories in this category.

Kolibri-1 (Aleph Alpha)

Aleph Alpha Kolibri-1, a 78B German-English MoE under Apache 2.0

Clef (Cloudflare)

Cloudflare Clef, open decision models

EmbeddingGemma 2 (Google DeepMind)

EmbeddingGemma 2, an open multimodal embedding model

d1-3B and d1-omni-600M (Liquid AI)

Liquid AI opens d1: d1-3B and d1-omni-600M decision models

pplx-embed-v2 (Perplexity)

Perplexity opens pplx-embed-v2 multimodal embedders

Holo4 (H Company)

H Company releases Holo4 computer-use models

Ling-3.0-flash-VL (Ant Group)

InclusionAI open-sources Ling-3.0-flash-VL: 124B vision-language MoE, 5.5B active, MIT

DeepSeek V4.1 Flash (DeepSeek)

DeepSeek V4.1 Flash: 552B MoE, 8B/16B active, KV cache 400x smaller than V1, MIT

Hy4 preview (Tencent)

Tencent Hy4 preview: 770B/49B Apache 2.0 MoE, Sherry quant shrinks 1.5 TB to 214 GB

Qwen3.8-Flash-Next (Alibaba Qwen)

Qwen3.8-Flash-Next previews the Qwen4 architecture in open weights

Apodex 1.1 + FrontierAgent (Apodex)

Apodex 1.1 agentic model family with open-weight 35B mini and FrontierAgent harness

GLM-5.3-Flash (Z.ai)

GLM-5.3-Flash: the OX Alpha mystery model, open-sourced under MIT

Qwen3.8-27B (Alibaba (Qwen))

Qwen3.8-27B ties GPT-5.6 Luna and runs on a 4090

Ling-3.0 (Ant Group (InclusionAI))

Ling-3.0: six open base checkpoints across training stages

LFM2.5 QAD checkpoints (Liquid AI)

Liquid AI ships LFM2.5 QAD 4-bit checkpoints for edge devices

Ornith-1.5 (Ornith)

Ornith-1.5 family: open 397B MoE matches Opus 4.8 on Terminal-Bench 2.1

YOLO26 (Ultralytics)

Ultralytics YOLO26 removes NMS from default inference entirely

dots3-note preview (Xiaohongshu (dots studio))

dots3-note preview: 280B MoE omni model with TEMPO RL for long-horizon agents

North Micro Vision (Cohere)

Cohere North Micro Vision: 2.4B VLM under Apache 2.0

DeepSeek V4 Pro 0813 (DeepSeek)

DeepSeek V4 Pro 0813 goes GA with MIT-licensed open weights

LFM2.5-VL-3B (Liquid AI)

Liquid AI LFM2.5-VL-3B runs 228 tok/s on M5 Max in ~3GB

Muse Glimmer 30B (Meta)

Meta returns to open source with Muse Glimmer 30B under Apache 2.0

Motif 3 (Motif Technologies)

Motif 3 from Korea: 314B MoE open-sourced under MIT

LFM2.5-2.6B (Liquid AI)

Liquid's LFM2.5-2.6B: agentic RL trained inside real harnesses, running in 1.7GB on a phone

LongCat-Flash-Lite-Sparse (Meituan (LongCat))

Meituan's LongCat-Flash-Lite-Sparse: 1M native context at 3B active parameters, MIT licensed

Inkling-Small (Thinking Machines)

Thinking Machines releases Inkling-Small: 276B/12B open MoE that beats its 975B sibling on agentic coding

Kimi K3 open weights (Moonshot AI)

Moonshot releases Kimi K3's full open-weight checkpoints — 2.8T parameters, the largest open model ever

Laguna S 2.1 (Poolside)

Poolside open-sources Laguna S 2.1, a 118B agentic coding model with a 1M-token context window

Kimi K3 (Moonshot AI)

Moonshot's Kimi K3 — 2.8T parameters — launches its API live mid-show, with full open weights following July 27

MOSS-VL-Realtime (OpenMOSS)

OpenMOSS open-sources MOSS-VL-Realtime, an 11B VLM that decides when to speak — and when to stay silent

Inkling (Thinking Machines)

Thinking Machines releases Inkling, a 975B open-weights MoE trained on 45T multimodal tokens

Bonsai 27B (PrismML)

PrismML compresses a full 27B model to 3.9GB so it runs on a phone

Robostral Navigate (Mistral AI)

Mistral releases Robostral Navigate, its first embodied-navigation model

Agents-A1 (Shanghai AI Lab)

Shanghai AI Lab releases Agents-A1, an Apache 2.0 agentic MoE

Kimi K2.7 Code (Moonshot AI)

Moonshot AI open-sources Kimi K2.7 Code for agentic coding

GLM-5.2 (Z.ai (Zhipu AI))

Z.ai releases GLM-5.2, a 753B open MoE with 1M context

GitHub: DietrichGebert/ponytail (⭐159,636)

Makes your AI agent think like the laziest senior dev in the room. The best code is the code you never wrote.

Gemma 4 12B (Google DeepMind)

Google drops Gemma 4 12B, an encoder-free multimodal local model

Holo 3.1 (H Company)

H Company launches Holo 3.1 local computer-use agent models

Mellum 2 (JetBrains)

JetBrains open-sources Mellum 2, a 12B MoE coding model

Nemotron 3 Ultra (NVIDIA)

NVIDIA releases Nemotron 3 Ultra, a 550B open-weight MoE for agents

MiniCPM5-1B (OpenBMB)

OpenBMB MiniCPM5-1B: new SOTA 1B open-weights model

Hy-MT2 (Tencent)

Tencent open-sources Hy-MT2 translation models under Apache 2.0

Command A+ (Cohere)

Cohere releases Command A+, a 218B Apache 2.0 MoE with 25B active params

Sapiens2 (Meta AI)

Meta Sapiens2: family of 6 human-centric vision models (0.1B-5B)

DeepSeek V4 (DeepSeek)

DeepSeek V4: 1.6T MoE with CSA+HCA attention and 1M context

Granite 4.1 (IBM)

IBM Granite 4.1: dense non-thinking models with top tool calling

Mistral Medium 3.5 (Mistral AI)

Mistral Medium 3.5: 128B dense flagship with 256K context

Nemotron 3 Nano Omni (NVIDIA)

NVIDIA Nemotron 3 Nano Omni: hybrid Transformer-Mamba MoE

SenseNova U1 (SenseTime)

SenseTime open-sources SenseNova U1 unified multimodal MoE

Talkie (Talkie (Alec Radford & David Duvenaud))

Talkie: 13B open-weight LLM trained only on pre-1930 text

GitHub: nexu-io/open-design (⭐100,216)

🎨 Best DeepSeek Harness Design Plugin. The open-source Claude Design alternative. 🖥️ Local-first desktop app.

Qwen3.6-27B (Alibaba (Qwen))

Qwen3.6-27B: dense Apache-2.0 model beats Alibaba's own 400B flagship

Kimi K2.6 (Moonshot AI)

Kimi K2.6: 1T MoE open-source SOTA on SWE-Bench Pro

Gemma 4 21B REAP (0xSero)

Gemma 4 21B REAP: 20% expert-pruned Gemma 4 26B MoE

Qwen 3.6-35B-A3B (Alibaba (Qwen))

Qwen 3.6-35B-A3B: Apache 2.0 MoE with 3B active hits 73.4% SWE-Verified

Super Gemma 4 26B Uncensored v2 (Jiunsong (@songjunkr))

Super Gemma 4 26B Uncensored v2 trends on HF with 0/100 refusals

Lyra 2.0 (NVIDIA)

NVIDIA Lyra 2.0: single image to explorable 3D worlds, Apache 2.0

GitHub: tt-a1i/archify (⭐81,179)

Turn any idea, plan, or codebase into a beautiful interactive diagram. An agent skill for Claude Code, Codex,

GLM-5.1 (Z.ai (Zhipu AI))

GLM-5.1 takes #1 open-source spot on SWE-Bench Pro at 58.4%

GitHub: JuliusBrussee/caveman (⭐110,776)

🪨 why use many token when few token do trick. Viral skill + proxy for coding agents that cuts 65% of tokens by

GitHub: career-ops-hq/career-ops (⭐73,909)

Open-source AI job search agent and job finder: scan job boards, score each job 1-5 against your CV before you

GitHub: Graphify-Labs/graphify (⭐125,033)

Turn any codebase, with its docs, SQL schemas, configs, and PDFs, into a queryable knowledge graph. A /graphif

Qwen3.5-Omni (Alibaba (Qwen))

Alibaba open-sources Qwen3.5-Omni, a 397B native omni-modal model

Gemma 4 (Google DeepMind)

Google releases Gemma 4 open-weights family under Apache 2.0

LFM2.5-350M (Liquid AI)

Liquid AI ships LFM2.5-350M with agentic tool calling at 350M params

Bonsai (PrismML)

PrismML releases Bonsai 1-bit models, an 8B model in 1.15 GB

Gemma 4 (Google)

Apache-2.0 open family (E2B, E4B, 26B MoE, 31B); 31B #3 open model on Arena.

GitHub: ultraworkers/claw-code (⭐194,972)

An agent-managed museum exhibit, built in Rust with Gajae-Code / LazyCodex — developed and maintained with no

GitHub: VoltAgent/awesome-design-md (⭐120,034)

A collection of DESIGN.md files analysis by popular brand design systems. Drop one into your project and let c

GitHub: calesthio/OpenMontage (⭐65,852)

World's first open-source, agentic video production system. 12 production pipelines, 100+ tools, 700+ agent sk

Reka Edge (Reka AI)

Reka AI ships Edge, a 7B multimodal VLM for sub-second on-device inference

Holotron-12B (H Company)

H Company's Holotron-12B: hybrid SSM computer-use model at 8.9k tok/s

Mistral Small 4 (Mistral AI)

Mistral Small 4: 119B MoE with 6B active unifies vision, coding, reasoning

GitHub: stablyai/orca (⭐88,606)

Orca is the ADE for working with a fleet of parallel agents. Run any coding agent with your own subscription.

GitHub: Egonex-AI/Understand-Anything (⭐85,759)

Graphs that teach > graphs that impress. Turn any code into an interactive knowledge graph you can explore, se

MiroThinker-1.7 (MiroMind)

MiroThinker-1.7 open-source research agent hits SOTA

Nemotron 3 Super 120B (NVIDIA)

NVIDIA releases Nemotron 3 Super 120B with $26B open-source bet

Covenant-72B (Templar)

Covenant-72B: a decentralized-trained open 72B LLM

GitHub: garrytan/gstack (⭐135,739)

Use Garry Tan's exact Claude Code setup: 23 opinionated tools that serve as CEO, Designer, Eng Manager, Releas

GitHub: karpathy/autoresearch (⭐97,590)

AI agents running research on single-GPU nanochat training automatically

Qwen3.5 Small Series (Alibaba (Qwen))

Alibaba releases Qwen3.5 small models (2B, 4B, 9B) for local use

Yuan 3.0 Ultra (IEIT (Yuan AI Lab))

Yuan AI Lab releases Yuan 3.0 Ultra open-weights model

Step 3.5 Flash Base (StepFun)

StepFun open-sources Step 3.5 Flash Base with its training stack

Qwen 3.5 (Alibaba (Qwen))

Qwen 3.5 lands: 35B/3B-active Medium outperforms the old 235B flagship

LFM2-24B-A2B (Liquid AI)

Liquid AI releases LFM2-24B-A2B, a laptop-friendly 24B MoE

GitHub: Panniantong/Agent-Reach (⭐94,889)

Give your AI agent eyes to see the entire internet. Read & search Twitter, Reddit, YouTube, GitHub, Bilibili,

Qwen3.5-397B-A17B (Alibaba (Qwen))

Alibaba opens Qwen 3.5: 397B-param multimodal MoE with only 17B active

Tiny Aya (Cohere Labs)

Cohere Labs releases Tiny Aya, a 3.35B multilingual model for 70+ languages

ZUNA (Zyphra)

Zyphra opens ZUNA, a 380M-param EEG brain-computer interface model

GitHub: Leonxlnx/taste-skill (⭐94,103)

Taste-Skill - gives your AI good taste. stops the AI from generating boring, generic slop

GitHub: diegosouzapw/OmniRoute (⭐74,676)

Never stop coding. Free MIT AI gateway: one endpoint, 359 providers (150+ free), 1200+ models Kimi, Claude, GP

MiniMax M-2.5 (MiniMax)

MiniMax M-2.5 hits 80.2% SWE-Bench Verified with 10B active params

GLM-5 (Zhipu AI (Z.ai))

Z.ai launches GLM-5, the open-weights agentic coding crown

Qwen3-Coder-Next (Alibaba (Qwen))

Qwen3-Coder-Next hits 70.6% SWE-Bench Verified with 3B active params

Intern-S1-Pro (InternLM (Shanghai AI Lab))

Intern-S1-Pro: 1 trillion parameter open MoE for scientific reasoning

Step 3.5 Flash (StepFun)

StepFun Step 3.5 Flash: frontier reasoning claims at 11B active params

GLM-OCR (Zhipu AI (Z.ai))

Z.ai GLM-OCR: 0.9B model takes #1 on OmniDocBench

Trinity Large (Arcee AI)

Arcee AI ships Trinity Large: 400B MOE trained in 33 days for $20M

Jan v3 (Jan AI)

Jan AI releases Jan v3, a 4B model built for fast local inference

Kimi K2.5 (Moonshot AI)

Moonshot AI releases Kimi K2.5, the new open-source king

GitHub: multica-ai/andrej-karpathy-skills (⭐217,746)

A single CLAUDE.md file to improve Claude Code behavior, derived from Andrej Karpathy's observations on LLM co

GitHub: mvanhorn/last30days-skill (⭐63,851)

AI agent skill that researches any topic across Reddit, X, YouTube, HN, Polymarket, and the web - then synthes

LFM2.5-1.2B-Thinking (Liquid AI)

Liquid AI's LFM2.5-1.2B-Thinking: on-device reasoning under 900MB

GLM-4.7-Flash (Z.AI (Zhipu))

GLM-4.7-Flash: 30B MoE local coding agent with only 3B active params

GitHub: rtk-ai/rtk (⭐82,788)

CLI proxy that reduces LLM token consumption by 60-90% on common dev commands. Single Rust binary, zero depend

GitHub: affaan-m/ECC (⭐275,970)

The agent harness performance optimization system. Skills, instincts, memory, security, and research-first dev

GitHub: colbymchenry/codegraph (⭐73,607)

Pre-indexed code knowledge graph, auto syncs on code changes, for Claude Code, Codex, Gemini, Cursor, OpenCode

LongCat Flash Thinking (Meituan (LongCat))

Meituan's LongCat Flash Thinking: 560B MoE with 27B active, MIT licensed

GitHub: ZhuLinsen/daily_stock_analysis (⭐66,099)

LLM 驱动的多市场股票智能分析系统:多源行情、实时新闻、决策看板与自动推送,支持零成本定时运行。 LLM-powered multi-market stock analysis system with multi-s

MiroThinker 1.5 (MiroMind AI)

MiroThinker 1.5: 30B search agent beats trillion-param models

NousCoder 14B (Nous Research)

NousCoder 14B: 7% LiveCodeBench jump in 4 days of RL training

Alpha Mayo (NVIDIA)

NVIDIA Alpha Mayo: open source reasoning self-driving models

Solar Open 100B (Upstage)

Upstage Solar Open 100B: 102B MoE trained on 19.7T tokens

GitHub: koala73/worldmonitor (⭐88,141)

Real-time global intelligence dashboard. AI-powered news aggregation, geopolitical monitoring, and infrastruct

GitHub: headroomlabs-ai/headroom (⭐74,841)

Compress tool outputs, logs, files, and RAG chunks before they reach the LLM. 20% fewer tokens for coding agen

DeepSeek R1 (DeepSeek)

DeepSeek R1: the open reasoning model that crashed NVIDIA's stock

MiniMax-01 (MiniMax (Hailuo))

MiniMax-01: open model with a 4M token context window

BOLMO (Allen AI)

Allen AI's BOLMO reaches byte-level parity with tokenized models

OLMO 2 (multimodal) (Allen AI)

Allen AI adds video-input multimodal OLMO models in 4B/7B/8B sizes

FunctionGemma (Google DeepMind)

FunctionGemma: Google's 270M function-calling model for edge agents

Nemotron 3 Nano (NVIDIA)

NVIDIA ships Nemotron 3 Nano, a 30B hybrid Mamba-MoE with full recipes

GitHub: gsd-build/get-shit-done (⭐64,330)

A light-weight and powerful meta-prompting, context engineering and spec-driven development system for Claude

Arcee Trinity (Arcee AI)

Arcee Trinity launches US-trained open MoE family

DeepSeek V3.2 / V3.2-Speciale (DeepSeek)

DeepSeek V3.2 and V3.2-Speciale post gold-medal reasoning under MIT license

Mistral 3 (Large 3 + Ministral 3) (Mistral AI)

Mistral returns to Apache 2.0 with Mistral Large 3 and Ministral 3

Hermes 4.3 (Nous Research)

Nous Research ships Hermes 4.3 36B with decentralized training

GitHub: code-yeongyu/oh-my-openagent (⭐69,919)

OmO: Just type "mass ulw" keyword with your prompt. Now you are the master of graph engineering.

GitHub: nextlevelbuilder/ui-ux-pro-max-skill (⭐134,234)

An AI skill that provides design intelligence for building professional UI/UX across multiple platforms.

DeepSeek Math V2 (DeepSeek)

DeepSeek Math V2: 685B open-weights model with IMO gold-level math

Fara-7B (Microsoft)

Microsoft ships Fara-7B, a 7B on-device computer use agent

INTELLECT-3 (Prime Intellect)

Prime Intellect releases INTELLECT-3, a 106B open MoE model

HunyuanOCR (Tencent (Hunyuan))

Tencent's 1B HunyuanOCR beats 72B models on OCRBench

GitHub: 666ghj/MiroFish (⭐77,471)

A Simple and Universal Swarm Intelligence Engine, Predicting Anything. 简洁通用的群体智能引擎,预测万物

GitHub: openclaw/openclaw (⭐391,530)

The AI that really does things. Any OS. Any Platform. The lobster way. 🦞

GitHub: pbakaus/impeccable (⭐79,013)

The design language that makes your AI harness better at design.

ERNIE-4.5-VL-28B-A3B-Thinking (Baidu)

Baidu open-sources ERNIE-4.5-VL-28B-A3B-Thinking visual reasoning model

VibeThinker-1.5B (WeiboAI)

WeiboAI releases VibeThinker-1.5B open reasoning model

OlmoEarth (Allen Institute for AI (Ai2))

Ai2 launches OlmoEarth foundation models and open Earth-intelligence platform

LongCat Flash Omni (Meituan (LongCat))

Meituan releases LongCat Flash Omni, a 560B (27B active) omni model

Kimi K2 Thinking (Moonshot AI)

Moonshot AI releases Kimi K2 Thinking, an open 1T-param reasoning MoE

GitHub: shanraisshan/claude-code-best-practice (⭐67,307)

from vibe coding to agentic engineering - practice makes claude perfect

Granite 4.0 Nano (IBM)

IBM Granite 4.0 Nano: ultra-efficient tiny models for edge deployment

Ming-flash-omni Preview (InclusionAI (Ant Group))

Ming-flash-omni Preview: sparse MoE omni-modal open model

MiniMax M2 (MiniMax)

MiniMax M2: open-source agentic model at 8% of Claude's price, 2x speed

Kimi Linear (Moonshot AI (Kimi))

Kimi Linear: 48B open model with linear attention and 1M context

Qwen3-VL 2B & 32B (Alibaba (Qwen))

Qwen3-VL adds compact 2B and 32B multimodal models

olmOCR 2 7B (Allen Institute for AI (Ai2))

Ai2 releases olmOCR 2 7B open OCR model

DeepSeek-OCR (DeepSeek)

DeepSeek-OCR turns text into compressed vision tokens for massive contexts

LFM2-VL-3B (Liquid AI)

Liquid AI ships LFM2-VL-3B tiny multilingual vision-language model

PokeeResearch-7B (Pokee AI)

PokeeResearch-7B: open-source SOTA deep research agent model

GitHub: ComposioHQ/awesome-claude-skills (⭐76,755)

A curated list of awesome Claude Skills, resources, and tools for customizing Claude AI workflows

Qwen3-VL 3B/8B (Alibaba (Qwen))

Qwen3-VL adds compact 3B and 8B open vision-language models

C2S-Scale 27B (Google DeepMind)

Google's C2S-Scale 27B validates a cancer hypothesis in living cells

KORMo 10B (KAIST)

KAIST releases KORMo, a bilingual Korean/English 10B open model

GitHub: msitarzewski/agency-agents (⭐158,534)

A complete AI agency at your fingertips - From frontend wizards to Reddit community ninjas, from whimsy inject