Kolibri-1 (Aleph Alpha)
Aleph Alpha Kolibri-1, a 78B German-English MoE under Apache 2.0
AI Year in Review · Oct 10, 2025 – Oct 10, 2026
182 stories in this category.
Aleph Alpha Kolibri-1, a 78B German-English MoE under Apache 2.0
Cloudflare Clef, open decision models
EmbeddingGemma 2, an open multimodal embedding model
Liquid AI opens d1: d1-3B and d1-omni-600M decision models
Perplexity opens pplx-embed-v2 multimodal embedders
H Company releases Holo4 computer-use models
Black Forest Labs open-sources Flux 3 Action, a 7B world action model
Xiaomi open-sources MiMo V2.6 Pro, top open-weights model on AA index
InclusionAI open-sources Ling-3.0-flash-VL: 124B vision-language MoE, 5.5B active, MIT
DeepSeek V4.1 Flash: 552B MoE, 8B/16B active, KV cache 400x smaller than V1, MIT
Tencent Hy4 preview: 770B/49B Apache 2.0 MoE, Sherry quant shrinks 1.5 TB to 214 GB
Qwen3.8-Flash-Next previews the Qwen4 architecture in open weights
Apodex 1.1 agentic model family with open-weight 35B mini and FrontierAgent harness
GLM-5.3-Flash: the OX Alpha mystery model, open-sourced under MIT
Qwen3.8-27B ties GPT-5.6 Luna and runs on a 4090
Ling-3.0: six open base checkpoints across training stages
Liquid AI ships LFM2.5 QAD 4-bit checkpoints for edge devices
Ornith-1.5 family: open 397B MoE matches Opus 4.8 on Terminal-Bench 2.1
Ultralytics YOLO26 removes NMS from default inference entirely
dots3-note preview: 280B MoE omni model with TEMPO RL for long-horizon agents
Cohere North Micro Vision: 2.4B VLM under Apache 2.0
DeepSeek V4 Pro 0813 goes GA with MIT-licensed open weights
Liquid AI LFM2.5-VL-3B runs 228 tok/s on M5 Max in ~3GB
Meta returns to open source with Muse Glimmer 30B under Apache 2.0
Motif 3 from Korea: 314B MoE open-sourced under MIT
DeepSeek Harness: Everything is a Plugin.
Liquid's LFM2.5-2.6B: agentic RL trained inside real harnesses, running in 1.7GB on a phone
Meituan's LongCat-Flash-Lite-Sparse: 1M native context at 3B active parameters, MIT licensed
Thinking Machines releases Inkling-Small: 276B/12B open MoE that beats its 975B sibling on agentic coding
Moonshot releases Kimi K3's full open-weight checkpoints — 2.8T parameters, the largest open model ever
Poolside open-sources Laguna S 2.1, a 118B agentic coding model with a 1M-token context window
Moonshot's Kimi K3 — 2.8T parameters — launches its API live mid-show, with full open weights following July 27
OpenMOSS open-sources MOSS-VL-Realtime, an 11B VLM that decides when to speak — and when to stay silent
Thinking Machines releases Inkling, a 975B open-weights MoE trained on 45T multimodal tokens
PrismML compresses a full 27B model to 3.9GB so it runs on a phone
Mistral releases Robostral Navigate, its first embodied-navigation model
Shanghai AI Lab releases Agents-A1, an Apache 2.0 agentic MoE
Meituan reveals LongCat-2.0, a 1.6T MoE trained entirely on Chinese ASICs
Moonshot AI open-sources Kimi K2.7 Code for agentic coding
Z.ai releases GLM-5.2, a 753B open MoE with 1M context
Makes your AI agent think like the laziest senior dev in the room. The best code is the code you never wrote.
Google drops Gemma 4 12B, an encoder-free multimodal local model
H Company launches Holo 3.1 local computer-use agent models
JetBrains open-sources Mellum 2, a 12B MoE coding model
NVIDIA releases Nemotron 3 Ultra, a 550B open-weight MoE for agents
Self-hosted AI workspace.
OpenBMB MiniCPM5-1B: new SOTA 1B open-weights model
Tencent open-sources Hy-MT2 translation models under Apache 2.0
Cohere releases Command A+, a 218B Apache 2.0 MoE with 25B active params
Meta Sapiens2: family of 6 human-centric vision models (0.1B-5B)
DeepSeek V4: 1.6T MoE with CSA+HCA attention and 1M context
IBM Granite 4.1: dense non-thinking models with top tool calling
Mistral Medium 3.5: 128B dense flagship with 256K context
NVIDIA Nemotron 3 Nano Omni: hybrid Transformer-Mamba MoE
SenseTime open-sources SenseNova U1 unified multimodal MoE
Talkie: 13B open-weight LLM trained only on pre-1930 text
🎨 Best DeepSeek Harness Design Plugin. The open-source Claude Design alternative. 🖥️ Local-first desktop app.
Qwen3.6-27B: dense Apache-2.0 model beats Alibaba's own 400B flagship
Kimi K2.6: 1T MoE open-source SOTA on SWE-Bench Pro
Gemma 4 21B REAP: 20% expert-pruned Gemma 4 26B MoE
Qwen 3.6-35B-A3B: Apache 2.0 MoE with 3B active hits 73.4% SWE-Verified
Super Gemma 4 26B Uncensored v2 trends on HF with 0/100 refusals
NVIDIA Lyra 2.0: single image to explorable 3D worlds, Apache 2.0
Tencent HYWorld 2.0 turns a single image into editable 3D scenes
Turn any idea, plan, or codebase into a beautiful interactive diagram. An agent skill for Claude Code, Codex,
Nous Research ships Hermes 27B, paired with the Hermes harness
GLM-5.1 takes #1 open-source spot on SWE-Bench Pro at 58.4%
🪨 why use many token when few token do trick. Viral skill + proxy for coding agents that cuts 65% of tokens by
Open-source AI job search agent and job finder: scan job boards, score each job 1-5 against your CV before you
Turn any codebase, with its docs, SQL schemas, configs, and PDFs, into a queryable knowledge graph. A /graphif
Alibaba open-sources Qwen3.5-Omni, a 397B native omni-modal model
Google releases Gemma 4 open-weights family under Apache 2.0
Liquid AI ships LFM2.5-350M with agentic tool calling at 350M params
PrismML releases Bonsai 1-bit models, an 8B model in 1.15 GB
Apache-2.0 open family (E2B, E4B, 26B MoE, 31B); 31B #3 open model on Arena.
An agent-managed museum exhibit, built in Rust with Gajae-Code / LazyCodex — developed and maintained with no
A collection of DESIGN.md files analysis by popular brand design systems. Drop one into your project and let c
World's first open-source, agentic video production system. 12 production pipelines, 100+ tools, 700+ agent sk
MiniMax 2.7 open-source weights discussed as small-model momentum continues
Reka AI ships Edge, a 7B multimodal VLM for sub-second on-device inference
H Company's Holotron-12B: hybrid SSM computer-use model at 8.9k tok/s
Mistral Small 4: 119B MoE with 6B active unifies vision, coding, reasoning
Learn it. Build it. Ship it for others.
Orca is the ADE for working with a fleet of parallel agents. Run any coding agent with your own subscription.
Graphs that teach > graphs that impress. Turn any code into an interactive knowledge graph you can explore, se
MiroThinker-1.7 open-source research agent hits SOTA
NVIDIA releases Nemotron 3 Super 120B with $26B open-source bet
Covenant-72B: a decentralized-trained open 72B LLM
Use Garry Tan's exact Claude Code setup: 23 opinionated tools that serve as CEO, Designer, Eng Manager, Releas
Write HTML. Render video. Built for agents.
AI agents running research on single-GPU nanochat training automatically
Alibaba releases Qwen3.5 small models (2B, 4B, 9B) for local use
Yuan AI Lab releases Yuan 3.0 Ultra open-weights model
StepFun open-sources Step 3.5 Flash Base with its training stack
The open-source app everyone uses to manage agents at work
Qwen 3.5 lands: 35B/3B-active Medium outperforms the old 235B flagship
Liquid AI releases LFM2-24B-A2B, a laptop-friendly 24B MoE
Perplexity launches pplx-embed SOTA embedding models
Give your AI agent eyes to see the entire internet. Read & search Twitter, Reddit, YouTube, GitHub, Bilibili,
Alibaba opens Qwen 3.5: 397B-param multimodal MoE with only 17B active
Cohere Labs releases Tiny Aya, a 3.35B multilingual model for 70+ languages
Zyphra opens ZUNA, a 380M-param EEG brain-computer interface model
Taste-Skill - gives your AI good taste. stops the AI from generating boring, generic slop
Production-grade engineering skills for AI coding agents.
Never stop coding. Free MIT AI gateway: one endpoint, 359 providers (150+ free), 1200+ models Kimi, Claude, GP
MiniMax M-2.5 hits 80.2% SWE-Bench Verified with 10B active params
Z.ai launches GLM-5, the open-weights agentic coding crown
Qwen3-Coder-Next hits 70.6% SWE-Bench Verified with 3B active params
Intern-S1-Pro: 1 trillion parameter open MoE for scientific reasoning
StepFun Step 3.5 Flash: frontier reasoning claims at 11B active params
Z.ai GLM-OCR: 0.9B model takes #1 on OmniDocBench
Skills for Real Engineers. Straight from my .agents directory.
Arcee AI ships Trinity Large: 400B MOE trained in 33 days for $20M
Jan AI releases Jan v3, a 4B model built for fast local inference
Moonshot AI releases Kimi K2.5, the new open-source king
A single CLAUDE.md file to improve Claude Code behavior, derived from Andrej Karpathy's observations on LLM co
AI agent skill that researches any topic across Reddit, X, YouTube, HN, Polymarket, and the web - then synthes
Liquid AI's LFM2.5-1.2B-Thinking: on-device reasoning under 900MB
GLM-4.7-Flash: 30B MoE local coding agent with only 3B active params
CLI proxy that reduces LLM token consumption by 60-90% on common dev commands. Single Rust binary, zero depend
The agent harness performance optimization system. Skills, instincts, memory, security, and research-first dev
Pre-indexed code knowledge graph, auto syncs on code changes, for Claude Code, Codex, Gemini, Cursor, OpenCode
M3: 235B open-source medical LLM claims to beat GPT 5.2 on HealthBench
Google releases MedGemma 1.5 for offline medical imaging
Meituan's LongCat Flash Thinking: 560B MoE with 27B active, MIT licensed
LLM 驱动的多市场股票智能分析系统:多源行情、实时新闻、决策看板与自动推送,支持零成本定时运行。 LLM-powered multi-market stock analysis system with multi-s
MiroThinker 1.5: 30B search agent beats trillion-param models
NousCoder 14B: 7% LiveCodeBench jump in 4 days of RL training
NVIDIA Alpha Mayo: open source reasoning self-driving models
Upstage Solar Open 100B: 102B MoE trained on 19.7T tokens
Real-time global intelligence dashboard. AI-powered news aggregation, geopolitical monitoring, and infrastruct
Compress tool outputs, logs, files, and RAG chunks before they reach the LLM. 20% fewer tokens for coding agen
Qwen 3 Coder posts insane scores in the race for the coding crown
DeepSeek R1: the open reasoning model that crashed NVIDIA's stock
DeepSeek V3.1 Terminus lands amid September's relentless pace
MiniMax-01: open model with a 4M token context window
Kimi K2: the Chinese open model that earned mainstream respect
Tencent enters the open weights race
GLM 4.5 runs on Cerebras fast enough to win hackathons
GLM 4.6 quietly becomes the model businesses actually use
Allen AI's BOLMO reaches byte-level parity with tokenized models
Allen AI adds video-input multimodal OLMO models in 4B/7B/8B sizes
FunctionGemma: Google's 270M function-calling model for edge agents
NVIDIA ships Nemotron 3 Nano, a 30B hybrid Mamba-MoE with full recipes
A light-weight and powerful meta-prompting, context engineering and spec-driven development system for Claude
Arcee Trinity launches US-trained open MoE family
DeepSeek V3.2 and V3.2-Speciale post gold-medal reasoning under MIT license
Mistral returns to Apache 2.0 with Mistral Large 3 and Ministral 3
Nous Research ships Hermes 4.3 36B with decentralized training
OmO: Just type "mass ulw" keyword with your prompt. Now you are the master of graph engineering.
An AI skill that provides design intelligence for building professional UI/UX across multiple platforms.
DeepSeek Math V2: 685B open-weights model with IMO gold-level math
Microsoft ships Fara-7B, a 7B on-device computer use agent
Prime Intellect releases INTELLECT-3, a 106B open MoE model
Tencent's 1B HunyuanOCR beats 72B models on OCRBench
A Simple and Universal Swarm Intelligence Engine, Predicting Anything. 简洁通用的群体智能引擎,预测万物
The AI that really does things. Any OS. Any Platform. The lobster way. 🦞
OLMo 3: Allen AI's fully open 32B model with complete recipe
Meta SAM 3: open-vocabulary segmentation and tracking in video
SAM 3D turns single photos into 3D objects and human bodies
The design language that makes your AI harness better at design.
Baidu open-sources ERNIE-4.5-VL-28B-A3B-Thinking visual reasoning model
H Company open-sources Holo2 multimodal computer-use agent family
WeiboAI releases VibeThinker-1.5B open reasoning model
Ai2 launches OlmoEarth foundation models and open Earth-intelligence platform
Meituan releases LongCat Flash Omni, a 560B (27B active) omni model
Moonshot AI releases Kimi K2 Thinking, an open 1T-param reasoning MoE
from vibe coding to agentic engineering - practice makes claude perfect
IBM Granite 4.0 Nano: ultra-efficient tiny models for edge deployment
Ming-flash-omni Preview: sparse MoE omni-modal open model
MiniMax M2: open-source agentic model at 8% of Claude's price, 2x speed
Kimi Linear: 48B open model with linear attention and 1M context
Qwen3-VL adds compact 2B and 32B multimodal models
Ai2 releases olmOCR 2 7B open OCR model
DeepSeek-OCR turns text into compressed vision tokens for massive contexts
Liquid AI ships LFM2-VL-3B tiny multilingual vision-language model
PokeeResearch-7B: open-source SOTA deep research agent model
A curated list of awesome Claude Skills, resources, and tools for customizing Claude AI workflows
Qwen3-VL adds compact 3B and 8B open vision-language models
Google's C2S-Scale 27B validates a cancer hypothesis in living cells
KAIST releases KORMo, a bilingual Korean/English 10B open model
A complete AI agency at your fingertips - From frontend wizards to Reddit community ninjas, from whimsy inject