<?xml version="1.0" encoding="UTF-8"?>
<rss version="2.0" xmlns:atom="http://www.w3.org/2005/Atom"><channel><title>Fresh Weights: new AI launches</title><link>https://freshweights.com/</link><description>Every new AI model, agent and dev tool worth knowing, ranked by real traction instead of hype. Updated four times a day, free.</description><language>en</language><atom:link href="https://freshweights.com/feed.xml" rel="self" type="application/rss+xml"/><item><title>NIRNAY 450M</title><link>https://freshweights.com/launch/nirnay-450m/</link><guid isPermaLink="true">https://freshweights.com/launch/nirnay-450m/</guid><pubDate>Thu, 01 Oct 2026 00:00:00 +0000</pubDate><category>models</category><description>Apache-2.0 local decision classifier built on Laya, returning probabilities for intent, choice, score and yes/no questions instead of generated text. Why it matters: Ships a downloadable checkpoint and evaluation files; author reports 87.9% on Banking77 after fine-tuning. Its Jev comparison is fine-tuned versus zero-shot, not a like-for-like model ranking; context is limited to 512 tokens.</description></item><item><title>Gemini Skills rollout</title><link>https://freshweights.com/launch/gemini-skills-rollout/</link><guid isPermaLink="true">https://freshweights.com/launch/gemini-skills-rollout/</guid><pubDate>Thu, 01 Oct 2026 00:00:00 +0000</pubDate><category>agents</category><description>Google is rolling reusable skills into Gemini chat, with slash-command invocation, reference files and multiple skills combined in one task. Why it matters: Skills will replace Gems through staged migrations; personal-account Gems are scheduled to lose support starting in November, while Workspace rollout and migration follow later.</description></item><item><title>Open Jarvis local-agent harness</title><link>https://freshweights.com/launch/open-jarvis-local-agent-harness/</link><guid isPermaLink="true">https://freshweights.com/launch/open-jarvis-local-agent-harness/</guid><pubDate>Thu, 01 Oct 2026 00:00:00 +0000</pubDate><category>agents</category><description>An open local-agent framework that makes model, inference engine, agent logic, tools and memory separately configurable. Why it matters: Uses a cloud teacher to propose harness changes and accepts improvements only after regression evaluation; Lambda reports recovering much of a small local model&#x27;s benchmark gap through harness retuning.</description></item><item><title>AG-UI 1.0</title><link>https://freshweights.com/launch/ag-ui-1-0/</link><guid isPermaLink="true">https://freshweights.com/launch/ag-ui-1-0/</guid><pubDate>Thu, 01 Oct 2026 00:00:00 +0000</pubDate><category>infra</category><description>Stable agent-to-application event protocol with schema-generated TypeScript, Python and .NET SDKs. Why it matters: Adds subagent events, multimodal tool results, approval interrupts, metadata and token usage while preserving compatibility with 0.x agents and clients.</description></item><item><title>Gemini 4 Argon</title><link>https://freshweights.com/launch/gemini-4-argon/</link><guid isPermaLink="true">https://freshweights.com/launch/gemini-4-argon/</guid><pubDate>Thu, 01 Oct 2026 00:00:00 +0000</pubDate><category>models</category><description>Google announced a frontier model for long-horizon coding, enterprise knowledge work and cybersecurity defense, initially restricted to trusted testers. Why it matters: Google reports a 1M-token output limit; access starts with Fairwind cyber defenders, with developer and consumer availability still pending.</description></item><item><title>GPT-Synopsys: OpenAI and Synopsys multi-year deal to build an AI model that operates EDA tools</title><link>https://freshweights.com/launch/gpt-synopsys-openai-and-synopsys-multi-year-deal-to-build-an-ai-model/</link><guid isPermaLink="true">https://freshweights.com/launch/gpt-synopsys-openai-and-synopsys-multi-year-deal-to-build-an-ai-model/</guid><pubDate>Thu, 01 Oct 2026 00:00:00 +0000</pubDate><category>models</category><description>OpenAI and Synopsys signed a multi-year agreement on Sep 30 to jointly develop GPT-Synopsys, a specialized model that directly operates Synopsys&#x27; EDA tools across chip design workflows, with revenue sharing and joint go-to-market. Why it matters: First frontier-lab partnership to put a model inside the semiconductor design loop: AI moves from advising engineers to running EDA tool flows, with sign-off verification kept as deterministic guardrails.</description></item><item><title>OpenAI Pro 500: $500/month tier with Ultrafast compute for GPT-6 Astra</title><link>https://freshweights.com/launch/openai-pro-500-500-month-tier-with-ultrafast-compute-for-gpt-6-astra/</link><guid isPermaLink="true">https://freshweights.com/launch/openai-pro-500-500-month-tier-with-ultrafast-compute-for-gpt-6-astra/</guid><pubDate>Thu, 01 Oct 2026 00:00:00 +0000</pubDate><category>infra</category><description>OpenAI announced a $500/month Pro 500 tier at DevDay (Sep 29) with &#x27;Ultrafast&#x27; compute for GPT-6 Astra in ChatGPT Work and 8x faster Codex; the $200 Pro plan reopens with halved compute allowance, and every Pro subscriber gets one Dot agent. Why it matters: OpenAI is now selling raw inference speed as a premium tier, with Ultrafast promising 8x faster Codex — a direct monetization of the price war with Anthropic as both eye IPOs.</description></item><item><title>Meshy crosses $100M ARR and launches official iOS/Android apps</title><link>https://freshweights.com/launch/meshy-crosses-100m-arr-and-launches-official-ios-android-apps/</link><guid isPermaLink="true">https://freshweights.com/launch/meshy-crosses-100m-arr-and-launches-official-ios-android-apps/</guid><pubDate>Thu, 01 Oct 2026 00:00:00 +0000</pubDate><category>media</category><description>Meshy announced Sep 30 that ARR surpassed $100M (up 100x in under two years, first AI-3D company at that milestone) and launched official iOS and Android apps for creating 3D models on phones. Why it matters: The first AI-3D company to hit $100M ARR — proof that text-to-3D is a real business, and mobile apps put 3D generation in 3D-printing hobbyists&#x27; pockets.</description></item><item><title>Claude for Government goes generally available to federal and state agencies</title><link>https://freshweights.com/launch/claude-for-government-goes-generally-available-to-federal-and-state/</link><guid isPermaLink="true">https://freshweights.com/launch/claude-for-government-goes-generally-available-to-federal-and-state/</guid><pubDate>Thu, 01 Oct 2026 00:00:00 +0000</pubDate><category>agents</category><description>Anthropic moved Claude for Government from July public beta to GA on Sep 30, delivering Claude coding and agentic capabilities in a FedRAMP High environment with audit logs, spend controls, and direct procurement. Why it matters: The first frontier agent platform to reach FedRAMP High GA — this is how agentic AI gets inside government workflows (memo drafting, RFP review, casework) with compliance baked in.</description></item><item><title>Naive-N0.5-Flash: 309B open-weights MoE with 1M context and no full-attention layers</title><link>https://freshweights.com/launch/naive-n0-5-flash-309b-open-weights-moe-with-1m-context-and-no-full/</link><guid isPermaLink="true">https://freshweights.com/launch/naive-n0-5-flash-309b-open-weights-moe-with-1m-context-and-no-full/</guid><pubDate>Thu, 01 Oct 2026 00:00:00 +0000</pubDate><category>models</category><description>Beijing startup NaiveAI released Naive-N0.5-Flash (Sep 27/28) under MIT: a 309B MoE (15.5B active) with native 1M-token context built from hybrid sliding-window + DeepSeek sparse attention, no full-attention layers anywhere, trained on Xiaomi MiMo-V2.5 base with AI-assisted research loops. Why it matters: The first large open-weights model with zero full-attention layers reaching 1M context — a real architectural bet on sparse attention at frontier scale, plus a live demo of AI-assisted model R&amp;D (151 optimization rounds in six days).</description></item><item><title>Mistral Forge: enterprise platform for training and continuously improving proprietary models</title><link>https://freshweights.com/launch/mistral-forge-enterprise-platform-for-training-and-continuously/</link><guid isPermaLink="true">https://freshweights.com/launch/mistral-forge-enterprise-platform-for-training-and-continuously/</guid><pubDate>Thu, 01 Oct 2026 00:00:00 +0000</pubDate><category>models</category><description>Mistral AI launched Forge (Sep 28) — a full-lifecycle model training platform (pre-training, SFT, DPO/ODPO, RL pipelines for policy alignment) letting enterprises and governments build and own models on their proprietary data; same week also shipped Mistral Small 4, open-source Leanstral code agent, and joined the Nvidia Nemotron Coalition. Why it matters: Mistral is pivoting from model releases to owning the enterprise training stack — a direct challenge to hyperscalers, betting companies want to own rather than rent their AI.</description></item><item><title>Magnitude: self-optimizing open-source inference engine for local agents</title><link>https://freshweights.com/launch/magnitude-self-optimizing-open-source-inference-engine-for-local-agents/</link><guid isPermaLink="true">https://freshweights.com/launch/magnitude-self-optimizing-open-source-inference-engine-for-local-agents/</guid><pubDate>Thu, 01 Oct 2026 00:00:00 +0000</pubDate><category>infra</category><description>Open-source inference engine (YC S25) that profiles your machine, recommends the best open models for it, and tunes kernels on-device; claims up to 2x faster than llama.cpp with dynamic memory allocation for parallel agent sessions. Why it matters: The first inference engine explicitly designed for local agents: long concurrent sessions get dynamic memory, model switching preserves tool use, and per-device kernel tuning reaches hardware-specific performance without hardware lock-in. 119 HN points / 54 comments on launch day.</description></item><item><title>tpu-megakernels (Inferact): open fused kernels that beat GB200 on Kimi K3</title><link>https://freshweights.com/launch/tpu-megakernels-inferact-open-fused-kernels-that-beat-gb200-on-kimi-k3/</link><guid isPermaLink="true">https://freshweights.com/launch/tpu-megakernels-inferact-open-fused-kernels-that-beat-gb200-on-kimi-k3/</guid><pubDate>Thu, 01 Oct 2026 00:00:00 +0000</pubDate><category>infra</category><description>Open-source fused &#x27;megakernels&#x27; for Google TPU v7 that serve Moonshot&#x27;s Kimi K3 at 709 tokens/sec (and Qwen3.8-27B at 1,515 tok/s) — ~57% faster than the same model on 16x NVIDIA GB200 GPUs, with unchanged accuracy. Why it matters: First credible public evidence that TPU inference software can beat NVIDIA&#x27;s flagship GB200 on open Chinese MoE models; one hand-fused Pallas kernel collapses hundreds of scheduled kernels, making TPU a first-class, cheaper serving target for Kimi K3/Qwen3.8-27B. Built by vLLM co-founders in partnership with Google Cloud.</description></item><item><title>Alibaba Zhenwu V900 AI accelerator</title><link>https://freshweights.com/launch/alibaba-zhenwu-v900-ai-accelerator/</link><guid isPermaLink="true">https://freshweights.com/launch/alibaba-zhenwu-v900-ai-accelerator/</guid><pubDate>Thu, 01 Oct 2026 00:00:00 +0000</pubDate><category>infra</category><description>Alibaba&#x27;s T-Head chip unit unveiled the Zhenwu V900 AI chip at the Apsara Conference, claiming 3x the performance of its M890 predecessor, 216GB HBM, and scale-out to 500,000-card clusters for frontier model training and inference. Why it matters: China&#x27;s most ambitious domestic AI-accelerator push yet — 500K-chip fabric aggregating 108PB of pooled memory is explicitly aimed at trillion-parameter Qwen MoEs; paired with a 20GW cloud buildout target and a 5-10T-parameter Qwen roadmap, it is the hardware backbone for NVIDIA-independent Chinese AI.</description></item><item><title>Robinhood Agents — in-app AI trading agents for retail investors</title><link>https://freshweights.com/launch/robinhood-agents/</link><guid isPermaLink="true">https://freshweights.com/launch/robinhood-agents/</guid><pubDate>Thu, 01 Oct 2026 00:00:00 +0000</pubDate><category>agents</category><description>Robinhood launched in-app AI agents at its HOOD Summit (Houston, Sep 29) that research markets, build strategies, and place trades within customer-set limits. Why it matters: First mainstream brokerage embedding end-to-end delegable trading agents directly in the core app — named agent, dedicated segregated account, configurable trade approvals.</description></item><item><title>CometAPI — unified OpenAI-compatible API for 500+ AI models</title><link>https://freshweights.com/launch/cometapi/</link><guid isPermaLink="true">https://freshweights.com/launch/cometapi/</guid><pubDate>Thu, 01 Oct 2026 00:00:00 +0000</pubDate><category>infra</category><description>A single API key and OpenAI-compatible interface to access 500+ frontier and multimodal AI models; launched Oct 1, 2026. Why it matters: Collapses per-provider integration/auth/SDK work into one endpoint at a moment when model choice is exploding faster than teams&#x27; integration capacity.</description></item><item><title>claude.dev — Anthropic&#x27;s developer publication hub for Claude builders</title><link>https://freshweights.com/launch/claude-dev/</link><guid isPermaLink="true">https://freshweights.com/launch/claude-dev/</guid><pubDate>Thu, 01 Oct 2026 00:00:00 +0000</pubDate><category>coding</category><description>Anthropic launched claude.dev on Sep 30, 2026 — one address for its engineering deep dives, Claude Code and API guides, agent playbooks, skills, and videos. Why it matters: Consolidates Anthropic&#x27;s own agent-building playbooks under one official, citable destination instead of scattered blog posts — a canonical reference for shipping Claude-powered products.</description></item><item><title>AIHOT — open-source framework for building automated industry-news sites</title><link>https://freshweights.com/launch/aihot/</link><guid isPermaLink="true">https://freshweights.com/launch/aihot/</guid><pubDate>Thu, 01 Oct 2026 00:00:00 +0000</pubDate><category>agents</category><description>Full-stack framework (Node 24 + PostgreSQL) that ingests RSS/webpages/X/WeChat, dedupes, double-scores, clusters with embeddings into hot topics, and publishes daily/weekly/monthly reports; 4,000+ GitHub stars within ~42 hours of release. Why it matters: A production-grade news-station framework, not a scraper — event heat scoring, a four-class relation evaluator, SelectBench evaluation tooling, and paid-call budget tracking, so swapping sources yields an industry news site for any vertical.</description></item><item><title>LiteLLM Lens — AI agents that analyze agent traces inside the gateway</title><link>https://freshweights.com/launch/litellm-lens/</link><guid isPermaLink="true">https://freshweights.com/launch/litellm-lens/</guid><pubDate>Thu, 01 Oct 2026 00:00:00 +0000</pubDate><category>infra</category><description>Trace-analysis product using AI agents to find recurring failures across agent runs: send OpenTelemetry traces to the LiteLLM proxy; Lens analyzes runs against team-defined criteria, groups similar issues, and links back to traces. Early access via waitlist. Why it matters: Turns the gateway from model-traffic routing into the agent-debugging chokepoint — traces analyzed where model calls already flow, without a separate hosted observability service; self-hosted (traces in ClickHouse, results in Postgres).</description></item><item><title>TeleOCR — China Telecom&#x27;s 1.2B open-source document parsing model</title><link>https://freshweights.com/launch/teleocr/</link><guid isPermaLink="true">https://freshweights.com/launch/teleocr/</guid><pubDate>Thu, 01 Oct 2026 00:00:00 +0000</pubDate><category>models</category><description>A ~1.2B open-source document parsing model from China Telecom&#x27;s Xingchen AGI Lab; unifies digital and camera-captured document parsing into structured outputs (Markdown, tables, formulas); HF trending with 30.4k downloads. Why it matters: A photographed contract parsed as accurately as a clean PDF — unifying digital + camera-captured documents in a lightweight 1.2B model.</description></item><item><title>Monid — persistent browser sessions so agents stay logged in across runs</title><link>https://freshweights.com/launch/monid/</link><guid isPermaLink="true">https://freshweights.com/launch/monid/</guid><pubDate>Thu, 01 Oct 2026 00:00:00 +0000</pubDate><category>agents</category><description>Monid (powered by TinyFish) lets AI agents reuse authenticated browser sessions across runs — sign in once, later runs start logged in (LinkedIn, Amazon, Workday examples); the password never reaches the agent. Why it matters: Solves the login-friction tax for browser agents — persistent session vaults make repeated logged-in workflows practical instead of re-authenticating every run.</description></item><item><title>Sparkling — Telegram-first trading app with an AI agent</title><link>https://freshweights.com/launch/sparkling/</link><guid isPermaLink="true">https://freshweights.com/launch/sparkling/</guid><pubDate>Thu, 01 Oct 2026 00:00:00 +0000</pubDate><category>agents</category><description>Trading app where an AI agent prepares every trade — Hyperliquid perps, Polymarket, tokenized US stocks (xStocks), cross-chain swaps — from one self-custodial balance; user confirms each pre-executed order card; gasless. Why it matters: &#x27;Agent prepares, human confirms&#x27; trading flow inside Telegram — brings Hyperliquid perps and prediction markets to chat-first users.</description></item><item><title>Jev-Omni — 12B multimodal decision classifier</title><link>https://freshweights.com/launch/jev-omni/</link><guid isPermaLink="true">https://freshweights.com/launch/jev-omni/</guid><pubDate>Thu, 01 Oct 2026 00:00:00 +0000</pubDate><category>models</category><description>12B multimodal decision classifier by akhilaaa3 — not a chat model. Built on Gemma 4 12B IT with a 30,000-question fine-tune and a classification head returning probabilities over 2–256 options for text, image, audio, and video. Apache-2.0. Why it matters: Jev-style typed decisions extended past text — one probability interface for text, image, audio, and video in a 12B open-weights model.</description></item><item><title>VisionHOPE — visual backbones as self-modifying learning systems</title><link>https://freshweights.com/launch/visionhope/</link><guid isPermaLink="true">https://freshweights.com/launch/visionhope/</guid><pubDate>Thu, 01 Oct 2026 00:00:00 +0000</pubDate><category>models</category><description>Official pretrained weights for &#x27;VisionHOPE: Visual Backbones as Self-Modifying Learning Systems&#x27; — hierarchical VisionHOPE-T/S/B checkpoints for ImageNet-1K classification, COCO detection/instance segmentation, ADE20K semantic segmentation. Why it matters: Research architecture framing visual backbones as self-modifying learning systems, released with T/S/B sizes across classification, detection, and segmentation.</description></item><item><title>Ling-3.1-flash — Ant Group&#x27;s 560B hybrid reasoning model</title><link>https://freshweights.com/launch/ling-3-1-flash/</link><guid isPermaLink="true">https://freshweights.com/launch/ling-3-1-flash/</guid><pubDate>Thu, 01 Oct 2026 00:00:00 +0000</pubDate><category>models</category><description>Ant Group&#x27;s InclusionAI launches a ~560B-total / ~25B-active hybrid reasoning MoE with up to 1M-token context, free for a two-week trial and promised open-source afterwards. Why it matters: The biggest open-weight-bound frontier-agentic model from a Chinese lab this quarter — targets coding, tool-use agents and long-horizon office/health workflows.</description></item><item><title>Index-Translate — Bilibili&#x27;s 150-language open translation family</title><link>https://freshweights.com/launch/index-translate/</link><guid isPermaLink="true">https://freshweights.com/launch/index-translate/</guid><pubDate>Thu, 01 Oct 2026 00:00:00 +0000</pubDate><category>models</category><description>Bilibili&#x27;s Index LLM team open-sources Apache-2.0 a translation model family — Index-Translate (150 text languages; 2B, 9B, 35B-A3B preview) plus Index-NativeLong (document translation), Index-Homura (syllable-budget dubbing), and Index-Echo (speech-to-speech with source voice). Why it matters: Purpose-built for real localization workflows — terminology/style control, cross-document consistency, dubbing scripts fitted to a syllable budget, voice-preserving speech translation — fully open under Apache-2.0.</description></item><item><title>Victoria + Maple — 44%-pruned Qwen3.8-Flash-Next and a Canada-first fine-tune</title><link>https://freshweights.com/launch/victoria-maple/</link><guid isPermaLink="true">https://freshweights.com/launch/victoria-maple/</guid><pubDate>Thu, 01 Oct 2026 00:00:00 +0000</pubDate><category>models</category><description>Community release by rmonsurate: Victoria is Qwen3.8-Flash-Next with 44% of experts cut (512→288/layer via REAP) and retrained at 4-bit (NVFP4) with a trained draft head — 70.0% Terminal-Bench 2.1 at 280 tok/s on a single NVIDIA B300; Maple is a Canadian-locale fine-tune on top. Why it matters: A rare community compression of a frontier MoE that keeps ~80% of agentic coding ability while fitting on one GPU with 2.08x speculative speedup — with reproducible evals and full training curves published.</description></item><item><title>Oído — Whisper-tiny-beating ASR on a $5 microcontroller</title><link>https://freshweights.com/launch/oido/</link><guid isPermaLink="true">https://freshweights.com/launch/oido/</guid><pubDate>Thu, 01 Oct 2026 00:00:00 +0000</pubDate><category>models</category><description>Open-source speech recognition from the Lokutor team: NVIDIA Conformer-CTC Small (13M params, int8) running on an ESP32-S3 with 8MB PSRAM — no GPU/NPU — beating Whisper-tiny.en on LibriSpeech and under real noise. Why it matters: Brings usable speech recognition to $5 hardware — two orders of magnitude cheaper than the laptop-class hardware Whisper-tiny needs — opening always-on voice interfaces for appliances, toys and remotes.</description></item><item><title>AI Phone — real-time translation inside phone calls</title><link>https://freshweights.com/launch/ai-phone/</link><guid isPermaLink="true">https://freshweights.com/launch/ai-phone/</guid><pubDate>Thu, 01 Oct 2026 00:00:00 +0000</pubDate><category>agents</category><description>Launched on Product Hunt Oct 1: an app that translates phone calls in real time across 100+ languages, with live transcripts, AI summaries and key-point highlights; claims 500K+ beta users and a 4.8/5 rating. Why it matters: Puts translation inside actual phone calls rather than a separate chat session — aimed at immigrants navigating deliveries, medical visits and fraud reporting — with transcript + summary capture.</description></item><item><title>CoreWeave Forge</title><link>https://freshweights.com/launch/coreweave-forge/</link><guid isPermaLink="true">https://freshweights.com/launch/coreweave-forge/</guid><pubDate>Wed, 30 Sep 2026 00:00:00 +0000</pubDate><category>infra</category><description>A connected AI development environment linking production serving, agent traces, data curation, model improvement and evaluation. Why it matters: Announced September 30 with new Agent Lens observability, model distillation and managed notebooks; ARIA and CPU/GPU sandboxes are now generally available.</description></item><item><title>Amazon Bedrock Managed Agents, powered by OpenAI (preview)</title><link>https://freshweights.com/launch/amazon-bedrock-managed-agents-powered-by-openai-preview/</link><guid isPermaLink="true">https://freshweights.com/launch/amazon-bedrock-managed-agents-powered-by-openai-preview/</guid><pubDate>Wed, 30 Sep 2026 00:00:00 +0000</pubDate><category>agents</category><description>AWS and OpenAI&#x27;s joint service lets enterprises build and run OpenAI-powered agents entirely inside their own AWS environment - IAM roles, CloudTrail auditing, preview in three US regions. Why it matters: OpenAI agents no longer require OpenAI&#x27;s infrastructure: orchestration and data stay inside the customer&#x27;s AWS account, which removes the biggest enterprise objection to OpenAI-based agents.</description></item><item><title>Visa open-sources VVAH: agentic vulnerability discovery and remediation harness</title><link>https://freshweights.com/launch/visa-open-sources-vvah-agentic-vulnerability-discovery-and-r/</link><guid isPermaLink="true">https://freshweights.com/launch/visa-open-sources-vvah-agentic-vulnerability-discovery-and-r/</guid><pubDate>Wed, 30 Sep 2026 00:00:00 +0000</pubDate><category>security</category><description>Visa Vulnerability Agentic Harness - a model-agnostic open framework whose agents discover, verify, remediate and validate software vulnerabilities through a multi-stage pipeline: attack-surface mapping, deep analysis, adversarial verification, then fixes. Why it matters: A payments giant publishing its internal cyber-defence agents (built from Anthropic&#x27;s Project Glasswing learnings) for inspection and adaptation - enterprise-grade offensive/defensive security as open code instead of a closed platform.</description></item><item><title>OpenAI Decisions API: low-latency deterministic choices via the Luna model</title><link>https://freshweights.com/launch/openai-decisions-api-low-latency-deterministic-choices-via-t/</link><guid isPermaLink="true">https://freshweights.com/launch/openai-decisions-api-low-latency-deterministic-choices-via-t/</guid><pubDate>Wed, 30 Sep 2026 00:00:00 +0000</pubDate><category>agents</category><description>Among 20+ DevDay updates: a Decisions API (limited preview) powered by the Luna model for low-latency, deterministic choices, plus major Codex upgrades (cloud tasks, workspaces) and a new $500/mo ChatGPT Pro tier with Astra Ultrafast. Why it matters: OpenAI is productizing the System-1/System-2 split the Jev/laya ecosystem pioneered - a dedicated fast-decision endpoint next to the reasoning models, which validates the category.</description></item><item><title>OpenClaw Enterprise: free, MIT-licensed control plane for persistent agents</title><link>https://freshweights.com/launch/openclaw-enterprise-free-mit-licensed-control-plane-for-pers/</link><guid isPermaLink="true">https://freshweights.com/launch/openclaw-enterprise-free-mit-licensed-control-plane-for-pers/</guid><pubDate>Wed, 30 Sep 2026 00:00:00 +0000</pubDate><category>infra</category><description>The open-source agentic framework goes enterprise: vendor-neutral control plane with multi-tenancy, hard security boundaries, lifecycle governance and auditing, while companies swap in their own models, harnesses and sandboxes. Originated inside OpenAI, now with the OpenClaw Foundation, backed by OpenAI, Red Hat and Nvidia. Why it matters: The agent-governance land grab just got a free, MIT-licensed contender with an unusual pedigree - centralized control over security, permissions and audit without model or sandbox lock-in.</description></item><item><title>MongoDB Atlas Agent Engine: execution, memory and governance for production agents</title><link>https://freshweights.com/launch/mongodb-atlas-agent-engine-execution-memory-and-governance-f/</link><guid isPermaLink="true">https://freshweights.com/launch/mongodb-atlas-agent-engine-execution-memory-and-governance-f/</guid><pubDate>Wed, 30 Sep 2026 00:00:00 +0000</pubDate><category>agents</category><description>Launched at MongoDB&#x27;s Investor Day: a unified execution, memory and governance layer for AI agents built into Atlas - accurate retrieval, persistent memory and enterprise security without stitching together a separate agent stack. Why it matters: A database incumbent&#x27;s answer to agent sprawl: agents run where the data already lives, model- and cloud-agnostic, instead of a fragile pile of tools that breaks on every framework update.</description></item><item><title>OpenAI DevDay: GPT-6.1 Sol and Dots always-on agents</title><link>https://freshweights.com/launch/openai-devday-gpt-6-1-sol-and-dots-always-on-agents/</link><guid isPermaLink="true">https://freshweights.com/launch/openai-devday-gpt-6-1-sol-and-dots-always-on-agents/</guid><pubDate>Wed, 30 Sep 2026 00:00:00 +0000</pubDate><category>models</category><description>Altman at DevDay: GPT-6.1 Sol, pitched as &#x27;very near-Astra-level intelligence at a fifth of the price&#x27; for coding and building, plus Dots - always-on enterprise agents powered by GPT-6 Astra, OpenAI&#x27;s answer to Meta&#x27;s agent push. Why it matters: Two moves at once: a cost-tier frontier model that undercuts its own flagship 5:1 for builders, and a persistent-agent product line aimed squarely at the enterprise rather than chat.</description></item><item><title>Landbase GTM-3 Omni: agentic GTM model that beats Clay, Apollo and ZoomInfo on list precision</title><link>https://freshweights.com/launch/landbase-gtm-3-omni-agentic-gtm-model-that-beats-clay-apollo-and/</link><guid isPermaLink="true">https://freshweights.com/launch/landbase-gtm-3-omni-agentic-gtm-model-that-beats-clay-apollo-and/</guid><pubDate>Wed, 30 Sep 2026 00:00:00 +0000</pubDate><category>agents</category><description>Next-gen model for agentic go-to-market discovery, qualification and outreach, shipped with a reproducible benchmark and native in Claude Code, Codex and Gemini CLI. Why it matters: First GTM model released alongside a reproducible, vendor-agnostic benchmark: 76.7% list precision vs 47.9% for Clay, 36.0% for Apollo and 23.2% for ZoomInfo, winning 18 of 26 prompts outright.</description></item><item><title>VoterCount.ai: plain-English questions over the US national voter file</title><link>https://freshweights.com/launch/votercount-ai-plain-english-questions-over-the-us-national-voter-file/</link><guid isPermaLink="true">https://freshweights.com/launch/votercount-ai-plain-english-questions-over-the-us-national-voter-file/</guid><pubDate>Wed, 30 Sep 2026 00:00:00 +0000</pubDate><category>agents</category><description>Conversational AI agent from L2 and IVC Media that answers plain-English questions about 217M+ registered voters, with a Professional-tier MCP server piping voter data into ChatGPT, Claude and other AI platforms. Why it matters: The only AI tool directly connected to a US national voter file: reporters, first-time candidates and researchers can get aggregate counts by geography, party, demographics and vote history in seconds, no data-vendor contract or analyst needed.</description></item><item><title>Yellow.ai Nexus EDGE: agentic desktop app that resolves employee IT, HR and Ops issues on the device</title><link>https://freshweights.com/launch/yellow-ai-nexus-edge-agentic-desktop-app-that-resolves-employee-it-hr/</link><guid isPermaLink="true">https://freshweights.com/launch/yellow-ai-nexus-edge-agentic-desktop-app-that-resolves-employee-it-hr/</guid><pubDate>Wed, 30 Sep 2026 00:00:00 +0000</pubDate><category>agents</category><description>Yellow.ai launched Nexus EDGE, an agentic desktop application that brings the Nexus Universal Agentic Interface to the employee&#x27;s computer. It sees IT, HR and Operations problems on the device, understands them against company knowledge and systems, and resolves them on the spot inside the user&#x27;s existing role-based permissions. Why it matters: First enterprise &#x27;Employee Experience Automation&#x27; play: it crosses from retrieval to resolution. A dead VPN before a deploy window, a failed new-hire install, a twice-a-year finance report get fixed on the device in real time, no knowledge-base article and no helpdesk queue.</description></item><item><title>VoiceStudio: open-source, fully-local ElevenLabs alternative with voice cloning, dubbing and dictation in 646 languages</title><link>https://freshweights.com/launch/voicestudio-open-source-fully-local-elevenlabs-alternative-with-voice/</link><guid isPermaLink="true">https://freshweights.com/launch/voicestudio-open-source-fully-local-elevenlabs-alternative-with-voice/</guid><pubDate>Wed, 30 Sep 2026 00:00:00 +0000</pubDate><category>models</category><description>VoiceStudio is an open-source, fully-local voice AI studio: voice cloning from a 3-second sample, voice design, video dubbing, dictation, transcription and audiobook creation in 646 languages. Ships via Docker with an official Google Colab notebook for zero-install use. Why it matters: Trending #1 on GitHub today with 49k stars: it shows local voice cloning and multilingual dubbing matching commercial quality without sending a byte of audio to a cloud API.</description></item><item><title>ZCode: Z.ai&#x27;s Apache-2.0 open-source coding-agent harness, open-sourced 2026-09-21</title><link>https://freshweights.com/launch/zcode-z-ai-s-apache-2-0-open-source-coding-agent-harness-open-sourced/</link><guid isPermaLink="true">https://freshweights.com/launch/zcode-z-ai-s-apache-2-0-open-source-coding-agent-harness-open-sourced/</guid><pubDate>Wed, 30 Sep 2026 00:00:00 +0000</pubDate><category>coding</category><description>Z.ai&#x27;s official coding-agent harness, an Electron desktop app plus browser interface, terminal agent and agent runtime built for GLM models, open-sourced under Apache-2.0 on 2026-09-21. The release came three days after users caught the desktop client silently packaging entire git histories (86.6% of the payload was .git) into encrypted archives for upload to Z.ai&#x27;s cloud. Why it matters: A rare look at a frontier lab&#x27;s first-party agent harness source, published under Apache-2.0 only because of a privacy scandal; it exposes the real client-side design of a Claude-Code-style system, with the upload pipeline scrubbed from the public code.</description></item><item><title>kev: open, trainable Jev-like family of small decision models on Qwen3.5/3.8</title><link>https://freshweights.com/launch/kev-open-trainable-jev-like-family-of-small-decision-models-on-qwen3-5/</link><guid isPermaLink="true">https://freshweights.com/launch/kev-open-trainable-jev-like-family-of-small-decision-models-on-qwen3-5/</guid><pubDate>Wed, 30 Sep 2026 00:00:00 +0000</pubDate><category>models</category><description>A family of small decision models (0.8B to 27B) built on Qwen3.5/3.8 that reproduces Jev&#x27;s prefill-only typed-decision architecture: yes/no, multiple-choice and rating questions answered as calibrated probabilities in one forward pass, with a TypeSafe System One-compatible API that works with the official Python SDK unchanged. Why it matters: The first open, trainable reproduction of a typed-decision architecture with published frozen evals: Kev-9B scores 0.837 out-of-domain accuracy on the locked test, and the Qwen3.5 port reportedly cost about $95 of H100 time.</description></item><item><title>OpenRig: open-source local control plane that runs Claude Code and Codex as one managed team</title><link>https://freshweights.com/launch/openrig-open-source-local-control-plane-that-runs-claude-code-and-codex/</link><guid isPermaLink="true">https://freshweights.com/launch/openrig-open-source-local-control-plane-that-runs-claude-code-and-codex/</guid><pubDate>Wed, 30 Sep 2026 00:00:00 +0000</pubDate><category>agents</category><description>A local control plane for multi-agent coding topologies: define agent teams (pods, seats, communication edges, continuity policies) in YAML, boot everything with `rig up`, and manage it live through CLI, TUI and an MCP server. The agents themselves are ordinary Claude Code and Codex sessions running in tmux. Why it matters: It manages the system the agents form, not the agents themselves: snapshot and restore whole topologies by name, broadcast messages across agents, adopt stray tmux sessions into a rig, and let agents manage their own topology via MCP tools.</description></item><item><title>DeepSeek open-sources Ascend infrastructure stack</title><link>https://freshweights.com/launch/deepseek-open-sources-ascend-infrastructure-stack/</link><guid isPermaLink="true">https://freshweights.com/launch/deepseek-open-sources-ascend-infrastructure-stack/</guid><pubDate>Wed, 30 Sep 2026 00:00:00 +0000</pubDate><category>infra</category><description>DeepSeek published Ascend-optimized versions of its NVIDIA-proven infra components — TileLang Ascend, DeepGEMM-Ascend, DeepEP-Ascend, FlashMLA Ascend, TileKernels Ascend backend, and DeepSelect — developed with Huawei&#x27;s support around the Ascend 950 128-card supernode. Why it matters: Ports the exact kernel stack that trained DeepSeek V4 (TileLang carrying most V4 training operators, DeepGEMM near hardware peak) 1:1 to Huawei Ascend NPUs, so any team can train and serve MoE models on domestic Chinese chips without rewriting Ascend C kernels by hand — a direct strike at CUDA lock-in.</description></item><item><title>Ponytail: the &#x27;lazy senior dev&#x27; skill for AI coding agents</title><link>https://freshweights.com/launch/ponytail-the-lazy-senior-dev-skill-for-ai-coding-agents/</link><guid isPermaLink="true">https://freshweights.com/launch/ponytail-the-lazy-senior-dev-skill-for-ai-coding-agents/</guid><pubDate>Wed, 30 Sep 2026 00:00:00 +0000</pubDate><category>coding</category><description>Open-source agent skill that makes coding agents write only the code a task needs — ~54% less code measured agentically, without dropping safety checks. Why it matters: An MIT-licensed skill that puts a YAGNI decision ladder in front of code generation, measured honestly on a real FastAPI+React repo with the same agent (54% mean LOC cut, 100% safety retention, per-task numbers published). At ~148k stars and trending again today, it is the most-adopted answer to AI code bloat.</description></item><item><title>Context Mode: context-window optimization for AI coding agents</title><link>https://freshweights.com/launch/context-mode-context-window-optimization-for-ai-coding-agents/</link><guid isPermaLink="true">https://freshweights.com/launch/context-mode-context-window-optimization-for-ai-coding-agents/</guid><pubDate>Wed, 30 Sep 2026 00:00:00 +0000</pubDate><category>infra</category><description>MCP server for AI coding agents that sandboxes verbose tool output and persists session memory, cutting context use by up to 98%. Why it matters: Attacks both sides of the context problem: 11 MCP tools keep raw tool output (56KB Playwright snapshots, 986KB repo research) out of the window, while SQLite+FTS5 session memory lets agents resume exactly after compaction. 24k stars, 17 platforms, trending again as context limits stay the bottleneck.</description></item><item><title>LiveNerf: a pre-registered benchmark for post-release model drift</title><link>https://freshweights.com/launch/livenerf-a-pre-registered-benchmark-for-post-release-model-drift/</link><guid isPermaLink="true">https://freshweights.com/launch/livenerf-a-pre-registered-benchmark-for-post-release-model-drift/</guid><pubDate>Wed, 30 Sep 2026 00:00:00 +0000</pubDate><category>security</category><description>Open-source, pre-registered 30-day benchmark that detects whether Claude Opus 5.5 quietly gets worse after launch. Why it matters: The first day-0-baseline instrument for the &#x27;was it nerfed?&#x27; question: frozen prompts, pinned CLI, exact graders, and a locked 78-question panel measured daily against launch-week baseline with clustered standard errors, on the UK AISI Inspect framework. Hit the HN front page today (694 points) because nobody has ever had a clean launch-day baseline before.</description></item><item><title>Pi.dev adds MCP to core: &#x27;You Said No MCP&#x27;</title><link>https://freshweights.com/launch/pi-dev-adds-mcp-to-core-you-said-no-mcp/</link><guid isPermaLink="true">https://freshweights.com/launch/pi-dev-adds-mcp-to-core-you-said-no-mcp/</guid><pubDate>Wed, 30 Sep 2026 00:00:00 +0000</pubDate><category>coding</category><description>Pi.dev reverses its public no-MCP stance, integrating Model Context Protocol into the core of its agentic coding platform with a &#x27;Codemode&#x27; sandbox for tool composition. Why it matters: A flagship agent platform admitting MCP won - plus its Codemode sandbox reframes MCP tools as composable OpenAPI-style endpoints, which could reshape how small harnesses consume tools.</description></item><item><title>xAI Team Bots</title><link>https://freshweights.com/launch/xai-team-bots/</link><guid isPermaLink="true">https://freshweights.com/launch/xai-team-bots/</guid><pubDate>Wed, 30 Sep 2026 00:00:00 +0000</pubDate><category>agents</category><description>Shared Grok Bots that work and learn alongside teams, built around a role or workflow with files, plugins, credentials, and memory. Why it matters: Turns personal AI agents into shared digital coworkers: one bot per team role with a Slack handle, private per-user conversations, and app plugins (Salesforce, Notion, GitHub).</description></item><item><title>Google DeepMind SynthID Bio</title><link>https://freshweights.com/launch/google-deepmind-synthid-bio/</link><guid isPermaLink="true">https://freshweights.com/launch/google-deepmind-synthid-bio/</guid><pubDate>Wed, 30 Sep 2026 00:00:00 +0000</pubDate><category>security</category><description>Watermarking technology that embeds an imperceptible, verifiable signature into AI-designed protein sequences and predicted 3D structures. Why it matters: First provenance layer for AI-designed biology: lab-tested watermarking that preserves biological function, aimed at biosecurity and stopping contaminated synthetic structures from polluting open databases.</description></item></channel></rss>
