ased; Anthropic Launc
nches Claude Marke
lace; Crusoe Raise
ases & Upda
r-1 — Fireworks Resea
lt on Kimi K3 th
news
AI Daily|Xiaomi MiMo-V2.6 Released; Anthropic Launches Claude Marketplace; Crusoe Raises $3.9B
2026-09-24
AI Daily|Xiaomi MiMo-V2.6 Released; Anthropic Launches Claude Marketplace; Crusoe Raises $3.9B
Model Releases & Updates
MiMo-V2.6 Pro & Flash — Xiaomi
- TL;DR: Xiaomi has open-sourced the MiMo-V2.6 Pro and Flash multimodal models, with the Pro variant achieving top-tier performance on the Artificial Analysis Intelligence Index and matching leading agent benchmarks.
- Key Highlights:
- Trained extensively via large-scale reinforcement learning, hitting an Intelligence Index score of 46 among open-weight models.
- Demonstrates agentic reasoning capabilities on par with proprietary frontier systems like Claude Opus 5 and GPT-5.6 Sol.
- Optimized for both heavy cloud orchestration and efficient edge deployments through its Flash variant.
- Specs: Open Weights / Multimodal / SOTA Open Agent Performance
- Links:
Never bet against open weight models https://t.co/XnNMiK3l9v
— Aravind Srinivas (@AravSrinivas) September 23, 2026
Ember-1 — Fireworks Research
- TL;DR: Fireworks Research has introduced Ember-1, a specialized model built on Kimi K3 that reduces reasoning token consumption by roughly 40% while preserving top-tier output quality.
- Key Highlights:
- Engineered to maximize token utility by generating shorter, highly streamlined reasoning traces.
- Available in research preview on Fireworks Serverless for developers seeking lower latency and cost.
- Significantly optimizes complex reasoning workflows without sacrificing accuracy.
- Specs: Based on Kimi K3 / ~40% Fewer Reasoning Tokens / Serverless API
- Links: Fireworks Research Blog
Nemotron 3 Diarization — NVIDIA
- TL;DR: NVIDIA has open-sourced Nemotron 3 Diarization, a lightweight 100M-parameter model designed to track up to eight speakers in live, overlapping conversations with high precision.
- Key Highlights:
- Secured the #1 ranking on initial Diarization-Bench evaluations with a 14.72% error rate (~24% lower than the runner-up).
- Seamlessly handles rapid one-second speech chunks and complex multi-speaker cross-talk.
- Features day-zero integration with Hugging Face Transformers under a commercial-friendly license.
- Specs: 100M Parameters / Up to 8 Speakers / Open Weights (Hugging Face)
- Links:
When several people talk at once, a transcript can get messy fast.
— NVIDIA AI (@NVIDIAAI) September 23, 2026
Our new Nemotron 3 Diarization model tracks who spoke when, even when voices overlap. It handles up to eight speakers, has 100M parameters, and is now available on @huggingface 🤗 pic.twitter.com/Mw9saKlN4E
Product Releases & Updates
Claude Marketplace — Anthropic
- What’s New: Anthropic has launched Claude Marketplace, a centralized hub for discovering and integrating third-party plugins, workspace connectors (such as Slack and Notion), enterprise agents, and professional service partners.
- Who It’s For: Developers, enterprise teams, and knowledge workers looking to scale their automated workflows.
- Try It: Anthropic Blog
Gemini Omni in Google Vids — Google Workspace
- What’s New: Google has integrated Gemini Omni into Google Vids, allowing any user with a Google account to produce high-resolution HD videos directly from their web browser using natural language prompts.
- Who It’s For: Content creators, marketers, educators, and general users seeking frictionless video generation.
- Try It: Google Workspace Blog
Recraft V4.1 Flash — Recraft AI
- What’s New: Recraft AI has rolled out V4.1 Flash, its fastest image generation model to date, producing high-end editorial and graphic compositions in approximately 1.3 seconds with zero workflow lag.
- Who It’s For: Designers, creative directors, and fast-paced product teams.
- Try It:
The fastest image model on the market is here.
— Recraft (@recraftai) September 23, 2026
Zero lag in your workflow.
Recraft V4.1 Flash produces images in 1.3 seconds.
Pick Flash from the model picker, generate, then hit Refine to clean up the details.
Check it out and give it a try ↓ pic.twitter.com/AJLvv7jv7j
Industry News
Crusoe Secures $3.9 Billion Funding at $30.9B Valuation
- What Happened: Data center builder and specialized cloud provider Crusoe closed a massive $3.9 billion financing round led by Atreides Management, Valor Equity Partners, and Mubadala, bringing its valuation to $30.9 billion.
- Why It Matters: Underscores the heavy private capital commitments flowing into specialized energy and data center infrastructure to meet escalating AI compute demand, amid broader debates over grid power constraints.
- Source: Newcomer Report
Stripe Details Token Economics & OpenRouter Acquisition Strategy
- What Happened: During its Stripe Tour flagship event, Stripe executives outlined the strategic driver behind its impending acquisition of OpenRouter, positioning payments and multi-model token routing as core infrastructure for the emerging agentic commerce economy.
- Source: WeChat Tech Report
Research Papers
WFM: Wiki Foundation Model for Agent Long-Term Memory — Research Team / DAIR.AI
- Motivation: AI agents managing long-term memory across complex, multi-step tasks often struggle with unstructured folder layouts and inefficient retrieval when navigating massive markdown stores.
- Key Innovation: Introduces WFM (Wiki Foundation Model), which treats agent memory as a linked markdown wiki graph, utilizing message passing conditioned on queries to fuse page text and link structure.
- Results: Achieves a 10.5x faster GPU-to-GPU training protocol and robust performance gains across five agent memory and multi-hop reasoning benchmarks.
- Paper: arXiv:2609.18182
Other Highlights
OpenRouter Batch API
- Overview: OpenRouter launched its new Batch API, enabling asynchronous batch processing across more than 70 supported models to optimize large-scale inference costs and throughput.
- Link: OpenRouter Announcement

