Live · rescanned 4× a day
Every AI launch worth knowing. Ranked by traction, not hype.
Fresh Weights tracks new models, agents and dev tools from GitHub, Hugging Face, X, Hacker News, Reddit and the Chinese labs. Every entry is scored on real traction, so the top of the list is what developers are actually picking up.
One email every Saturday: the week's top launches and 3 product ideas you could build with them. Free, unsubscribe anytime.
- –launches this week
- –added last 24h
- –build ideas
- 30+lab accounts watched
★ Spotlight
The 4 launches worth your time this week. Picked once a day by five AI models (GPT-6 Sol, Claude Sonnet 5.5, Xiaomi MiMo, Qwen and Kimi) that argue over the week's launches, with JEV breaking ties. The weekly newsletter shows the same picks.
Spotlight
- mattpocock/skills: Skills for Real Engineers: Matt Pocock's open collection of production agent skills, published straight from his .agents directory - 272k stars and #8 on GitHub…
- VoiceStudio: open-source, fully-local ElevenLabs alternative with voice cloning, dubbing and dictation in 646 languages: VoiceStudio is an open-source, fully-local voice AI studio: voice cloning from a 3-second sample, voice design, video dubbing, dictation…
- Paperclip: the app people use to manage agents at work: Open-source orchestration for teams of AI agents: org charts, budgets, governance, heartbeats and a ticket system in one dashboard.
- Pi.dev adds MCP to core: 'You Said No MCP': Pi.dev reverses its public no-MCP stance, integrating Model Context Protocol into the core of its agentic coding platform with a…
Latest AI launches
- Strands Decider 2B: AWS's open-source 2B-parameter decision model that picks among predefined options with calibrated confidence scores, returning answers in…
- Shopify Canvas: Shopify's new chat-driven store builder: merchants describe changes to the Sidekick AI agent and watch real theme code update live.
- Yandex Sona: Yandex's generative recommendation model that replaces a full multi-stage ranking pipeline with a single model, with no hand-engineered…
- Ivo Sage: Ivo's open-source model post-trained for long-horizon contract work, built by RL-fine-tuning DeepSeek V4 Flash on attorney-generated data…
- Weave Router 2.0: Open-source routing model that plugs into coding agents and dispatches each task to the best LLM, matching GPT-6 Astra's coding-agent…
- Janus: Single Go binary that runs GGUF models locally via llama.cpp with Vulkan (AMD/Intel/NVIDIA) or CPU fallback, exposing an OpenAI-compatible…
- DeepSeek dsh-libreoffice-kit: Open-source Node.js kit from DeepSeek that converts and renders Office documents (Word/PowerPoint/Excel to PDF and PNG) using prebuilt…
- Microsoft releases 301,000 Copilot coding-agent traces: Microsoft open-sourced 301,026 GitHub Copilot coding-agent sessions (9.3M model calls, 8.7M tool calls) with timings, token counts, cache…
- Microsoft MAI-Transcribe-2-Streaming + MAI-Voice-2.1: Microsoft launched three in-house voice models: MAI-Transcribe-2-Streaming (real-time transcription, 60 languages, first hypotheses in…
- Ai2 Olmo-Core 3: open MoE training stack: The Allen Institute for AI released Olmo-Core 3, a redesigned open training framework for mixture-of-experts models, benchmarked to 1.2T…
- Cohere Embed 5 (Pro + Fast): Cohere released Embed 5, a multimodal embedding family in two tiers — Pro ($0.12/1M) for max retrieval quality and Fast ($0.08/1M) for…
- Perplexity pplx-embed-v2-context-9b-preview: Perplexity Research and turbopuffer released a 9B MIT-licensed contextual embedding model for RAG that encodes each chunk with the full…
- LlamaIndex Extract v2.5: LlamaIndex launched Extract v2.5, a new generation of schema-based document extraction agents with a purpose-built agent harness…
- Grok for Intune: SpaceXAI released Grok for Intune, a separate enterprise iOS app — the Microsoft Intune edition of Grok with MAM support, app protection…
- MiniMax OpenAgentCore: MiniMax open-sourced OpenAgentCore (MIT), a self-hosted implementation of the OpenAI Agents API with native Codex, Claude Code and MiniMax…
- Suno Speech (Beta): Suno's new model generates spoken audio with matching background music in a single take — bedtime stories over soft piano, hype speeches…
- Perplexity pplx-decider-v1-27b: Open-weight 27B decision model fine-tuned from Qwen3.8-27B for classification and judgment tasks, plus a new Decisions API at $0.04 per…
- TileLang: Pythonic domain-specific language on top of TVM for writing high-performance GPU/CPU kernels — GEMM, dequant GEMM, FlashAttention…
- Superpowers: Complete software development methodology for coding agents: a set of composable skills (TDD, systematic debugging, brainstorming, plan…
- Impeccable: Design skill for AI coding agents: 1 skill, 24 commands, live browser iteration, and 61 deterministic detector rules that catch…
- Cursor Plugins (official marketplace): Cursor's official plugin specification and marketplace repo: 12 agent plugins including Ralph Loop self-iteration, Thermos branch review…
- MiMo-V2.6 MOPD2 checkpoints (Flash + Pro): Xiaomi MiMo released MOPD2-upgraded MiMo-V2.6-Flash and Pro checkpoints: a multi-teacher on-policy distillation stage that fuses…
- ExplorationBench: AI exploration in verifiable alien worlds: Tencent's Hunyuan team with Fudan and Tsinghua released ExplorationBench, a benchmark measuring whether AI systems can discover hidden…
- Maxwell: embodied AI model tops Meta-World physical-task benchmark: Researchers at the Chinese Academy of Sciences' Institute of Artificial Intelligence for Industries reported Maxwell, an embodied AI model…
- MatVerse: first 10B-parameter native-multimodal materials science model: Suzhou National Lab, Shanghai AI Lab and Shanghai Jiao Tong University unveiled MatVerse, a multimodal foundation model for materials…
- DeepSeek-V4.1-Flash: 552B MoE with 890-byte-per-token KV cache: DeepSeek's MIT-licensed multimodal MoE uses a causal encoder-decoder (CED) design with cross-layer KV sharing, shrinking the KV cache to…
- yomiyasu: agent skill that deodorizes AI-generated Japanese: MIT-licensed agent skill by ALGO ARTIS that rewrites AI-generated Japanese into natural, information-dense technical Japanese via 7…
- WeftCut: open-source video editor driven by AI agents over MCP: Free MIT-licensed desktop NLE (macOS/Windows/Linux) that exposes its entire timeline — place, trim, split, keyframe, caption, checkpoint…
- Boston Dynamics fits Atlas humanoid with a factory-ready four-fingered hand: Atlas's new hand has 13 degrees of freedom (nearly double the prior three-fingered generation) with pressure sensors across fingertips and…
- MixVio AI: unified creative workspace across 26 AI generation routes: New workspace unites video, image, and audio generation under one account, credit system, and private asset library — 14 model families…
- Tavus Griffin: the first Human Interaction Model: Tavus launched Griffin, a full-duplex video-to-video model that listens, watches, and generates expressive speech and video in real time…
- OpenAI disrupts Moonshot-linked 'adversarial distillation' campaign: OpenAI disclosed it shut down a coordinated campaign by 15,000+ users to extract hidden model reasoning, attributing a core cluster to…
- Inception Mercury Decide: structured decision model via the System One Decisions API: Inception AI released Mercury Decide, a structured decision model that takes a state plus typed yes/no, choice, and score questions and…
- Cantina Apex Flash-1: open-weight security investigation model: Cantina (co-founded by Spearbit's Harikrishnan Mulackal) released Apex Flash-1 on Oct 1, 2026: an open-weight 321B model fine-tuned from…
- Neuron AI v4: PHP agent framework release adds background streaming channels, frontend-executed tools, semantic memory and a revised workflow API.
- JetBrains Air in IDEs EAP: JetBrains opens early access to its multi-agent IDE experience through a Marketplace plugin and 2026.3 EAP builds.
- Cloudflare AI Search GA: Managed indexing and retrieval pipeline reaches general availability with native image embeddings, scanned-PDF OCR and files up to 10 MiB.
- NIRNAY 450M: Apache-2.0 local decision classifier built on Laya, returning probabilities for intent, choice, score and yes/no questions instead of…
- Gemini Skills rollout: Google is rolling reusable skills into Gemini chat, with slash-command invocation, reference files and multiple skills combined in one task.
- Open Jarvis local-agent harness: An open local-agent framework that makes model, inference engine, agent logic, tools and memory separately configurable.
- AG-UI 1.0: Stable agent-to-application event protocol with schema-generated TypeScript, Python and .NET SDKs.
- Gemini 4 Argon: Google announced a frontier model for long-horizon coding, enterprise knowledge work and cybersecurity defense, initially restricted to…
- GPT-Synopsys: OpenAI and Synopsys multi-year deal to build an AI model that operates EDA tools: OpenAI and Synopsys signed a multi-year agreement on Sep 30 to jointly develop GPT-Synopsys, a specialized model that directly operates…
- OpenAI Pro 500: $500/month tier with Ultrafast compute for GPT-6 Astra: OpenAI announced a $500/month Pro 500 tier at DevDay (Sep 29) with 'Ultrafast' compute for GPT-6 Astra in ChatGPT Work and 8x faster…
- Meshy crosses $100M ARR and launches official iOS/Android apps: Meshy announced Sep 30 that ARR surpassed $100M (up 100x in under two years, first AI-3D company at that milestone) and launched official…
- Claude for Government goes generally available to federal and state agencies: Anthropic moved Claude for Government from July public beta to GA on Sep 30, delivering Claude coding and agentic capabilities in a…
- Naive-N0.5-Flash: 309B open-weights MoE with 1M context and no full-attention layers: Beijing startup NaiveAI released Naive-N0.5-Flash (Sep 27/28) under MIT: a 309B MoE (15.5B active) with native 1M-token context built from…
- Mistral Forge: enterprise platform for training and continuously improving proprietary models: Mistral AI launched Forge (Sep 28) — a full-lifecycle model training platform (pre-training, SFT, DPO/ODPO, RL pipelines for policy…
- Magnitude: self-optimizing open-source inference engine for local agents: Open-source inference engine (YC S25) that profiles your machine, recommends the best open models for it, and tunes kernels on-device…
- tpu-megakernels (Inferact): open fused kernels that beat GB200 on Kimi K3: Open-source fused 'megakernels' for Google TPU v7 that serve Moonshot's Kimi K3 at 709 tokens/sec (and Qwen3.8-27B at 1,515 tok/s) — ~57%…
- Alibaba Zhenwu V900 AI accelerator: Alibaba's T-Head chip unit unveiled the Zhenwu V900 AI chip at the Apsara Conference, claiming 3x the performance of its M890 predecessor…
- Robinhood Agents — in-app AI trading agents for retail investors: Robinhood launched in-app AI agents at its HOOD Summit (Houston, Sep 29) that research markets, build strategies, and place trades within…
- CometAPI — unified OpenAI-compatible API for 500+ AI models: A single API key and OpenAI-compatible interface to access 500+ frontier and multimodal AI models; launched Oct 1, 2026.
- claude.dev — Anthropic's developer publication hub for Claude builders: Anthropic launched claude.dev on Sep 30, 2026 — one address for its engineering deep dives, Claude Code and API guides, agent playbooks…
- AIHOT — open-source framework for building automated industry-news sites: Full-stack framework (Node 24 + PostgreSQL) that ingests RSS/webpages/X/WeChat, dedupes, double-scores, clusters with embeddings into hot…
- LiteLLM Lens — AI agents that analyze agent traces inside the gateway: Trace-analysis product using AI agents to find recurring failures across agent runs: send OpenTelemetry traces to the LiteLLM proxy; Lens…
- TeleOCR — China Telecom's 1.2B open-source document parsing model: A ~1.2B open-source document parsing model from China Telecom's Xingchen AGI Lab; unifies digital and camera-captured document parsing…
- Monid — persistent browser sessions so agents stay logged in across runs: Monid (powered by TinyFish) lets AI agents reuse authenticated browser sessions across runs — sign in once, later runs start logged in…
- Sparkling — Telegram-first trading app with an AI agent: Trading app where an AI agent prepares every trade — Hyperliquid perps, Polymarket, tokenized US stocks (xStocks), cross-chain swaps…
- Jev-Omni — 12B multimodal decision classifier: 12B multimodal decision classifier by akhilaaa3 — not a chat model. Built on Gemma 4 12B IT with a 30,000-question fine-tune and a…
Ideas worth building
- CallSentry: every customer call compliance-scored in real time
- Airgap Extract: audit-ready document intelligence that never leaves the building
- TraceRoute: the coding-agent router learned from production traces
- CounselStack: enterprise-grade contract review for the mid-market at one-tenth the seat cost
- Agent Ledger: the audit-and-approve console for agentic trading accounts
- Parity Router: a hardware-aware inference gateway that cuts open-model bills 40-60% with certified accuracy parity
- AuthorityFast: the continuous compliance pipeline that turns AI vendor authorization into a 90-day rolling process
- DriftGuard: pager alerts when your pinned model silently gets worse
- Gatehouse: deterministic pre-execution decision gates for agent actions
- Autopsy: every agent incident ships a root-cause report and a regression test
- ResolveLoop: the closed-loop quality layer that makes AI support bots resolve instead of deflect
- VendorPassport: runtime attestation that gets agent vendors through enterprise procurement
- ClickSure: portal automations that can't hallucinate clicks
- RuleGate — pre-trade discipline copilot for intraday index traders
- MergeQueue Triage — Jev-powered PR triage gate for agent-flooded repos
- StudioWhatsApp — local-GPU product photo studio for marketplace sellers
- ReconPilot — GST reconciliation copilot for Indian SMBs
The week in AI launches, ranked. Plus 3 ideas worth building.
Every Saturday. Five AI models (GPT-6 Sol, Claude Sonnet 5.5, Xiaomi MiMo, Qwen and Kimi) argue over the week's radar data, and JEV judges which ideas hold up. You get the 3 winners, each with a build plan, who pays and how it makes money.
- This week's Spotlight launches, with traction and real builds
- 3 product ideas with a build plan, buyers and pricing
- First Sunday of the month: the single best idea of the month, in depth
No spam, no sponsored rankings. One click to unsubscribe.
How the radar works
Wide net, primary sources
GitHub Trending, Hugging Face, ModelScope, Hacker News, Product Hunt, Reddit, X and Chinese tech media. Official lab announcements count from the minute they're posted.
JEV traction, not vibes
JEV is the decision model that reads each launch's evidence (stars, builds, mentions) and returns a traction score from 0 (obscure) to 1 (everywhere). It never sees vendor copy.
What people actually built
Community builds, implementation repos and a one-line usability verdict drawn from evidence, not vendor claims.
Ideas you can ship
Daily ideas that combine two or more launches, plus weekly ideas that five AI models debate and JEV judges.