news

AI Daily|Google Launches Gemini 3.8 Live with Extended Thinking; Salesforce Unveils Koa Enterprise Agent Model

September 18, 2026
Updated Sep 18
6 min read
amp
ases & Upda
gemini
dates Gemini 3.8 L
google
ing — Google DeepM
deepmind
oogle DeepMind TL;DR
inference
ering inference token
stepfun
ime — Stepfun TL;DR
news
AI Daily|Google Launches Gemini 3.8 Live with Extended Thinking; Salesforce Unveils Koa Enterprise Agent Model
2026-09-18

AI Daily | 2026-09-18

💡 This report is generated automatically and updates every morning at 9am (Taipei time).


Model Releases & Updates

Gemini 3.8 Live & Extended Thinking — Google DeepMind

  • TL;DR: Google DeepMind released Gemini 3.8 Live and Extended Thinking, introducing real-time speech-to-speech reasoning with synchronous progress reporting and asynchronous background tool execution.
  • Key Highlights:
    • The Extended Thinking version achieves 82.6 on Speech-to-Speech benchmarks, pairing multi-step logical planning with natural conversational fillers (“let me check…”) while executing background tasks.
    • Supports multimodal real-time interactions (simultaneous vision and voice) across 97 languages while significantly lowering inference token consumption and latency.
  • Specs: Multimodal Real-Time Speech Architecture / Asynchronous Tool-Calling / 97 Language Support
  • Links: Google DeepMind Blog |

Koa Enterprise Agent Model — Salesforce AI Research

  • TL;DR: Salesforce unveiled Koa, a custom enterprise agent model trained by expanding declarative Agent Script specifications into multi-turn reinforcement learning environments using Group Relative Policy Optimization (GRPO).
  • Key Highlights:
    • Built upon the open-weight Nemotron-3-Super-120B base architecture, Koa is trained directly on structured workflow configurations and simulated user personas.
    • Scores 69.41 on Tau2Bench (outperforming base models and GPT-4.1) and achieves 0.86 on CRM Bench with noticeably enhanced function-calling precision.
  • Specs: Open Weights / Nemotron-3-Super-120B Base / GRPO Post-Training / Declarative Workflow Optimization
  • Links: ArXiv Paper |

StepAudio 3 Realtime — Stepfun

  • TL;DR: Stepfun open-sourced the StepAudio 3 family, a suite of five on-device speech recognition and synthesis models optimized for edge hardware and low-latency local deployment.
  • Key Highlights:
    • Delivers flexible model sizing ranging from 0.1B to 3B parameters for ASR and up to 0.6B for TTS.
    • Enables fully offline, low-memory voice interactions on smartphones and local PCs.
  • Specs: Open Weights / On-Device Speech Models (0.1B–3B) / Local Inference Ready
  • Links: Stepfun Blog |

Product Releases & Updates

Claude for Financial Advisors & Salesforce Integration — Anthropic

  • What’s New: Anthropic significantly expanded Claude’s enterprise utility by releasing dedicated workflow connectors for financial institutions (Charles Schwab, BlackRock, Addepar, Envestnet) alongside an embedded Salesforce integration featuring 37 pre-built sales skills. Advisors and sales teams can now execute client meeting prep, pipeline reviews, and compliance documentation directly within conversational threads under human oversight.
  • Who It’s For: Financial advisors, registered investment advisors (RIAs), and enterprise sales operators.
  • Try It: Anthropic Blog |

Grok Build Persistent Memory & Guides — xAI

  • What’s New: xAI published comprehensive engineering guides and workflows for Grok Build, formalizing its persistent background workspace memory that logs architectural project decisions, conventions, and facts across sessions using /memory and /dream.
  • Who It’s For: Software engineers and multi-agent pipeline developers.
  • Try It: xAI Guides |

Page Shield Client-Side Security ML — Cloudflare

  • What’s New: Cloudflare deployed its automated Page Shield machine learning model into live traffic, successfully detecting advanced client-side JavaScript skimming campaigns that went entirely undetected by VirusTotal and standard security scanners.
  • Who It’s For: Web storefront owners, e-commerce security teams, and enterprise IT operators.
  • Try It: Cloudflare Blog

Industry News

Databricks Deploys OpenAI Astra to All 3,500 Engineers

  • What Happened: Databricks officially announced the company-wide deployment of OpenAI’s Astra model to its ~3,500 engineers, following a successful 200-person pilot where Astra demonstrated clear superiority over previous frontier models in high-level system design and complex code refactoring.
  • Why It Matters: Reflects rapid enterprise adoption of advanced agentic models for core infrastructure engineering despite higher token consumption and cost overheads.
  • Source:

Mozilla Releases 91-Page Open-Weights Ecosystem Report

  • What Happened: Mozilla published a comprehensive 91-page report analyzing the open-weights AI landscape, revealing that open models lag frontier proprietary systems by only ~4 months. While open models command nearly 20% of OpenRouter traffic volume (led heavily by Chinese open-source releases), they capture just 4% of total model-layer revenue due to pricing and capability gaps.
  • Why It Matters: Highlights the widening wedge between open-source adoption breadth and direct monetization in the foundational model market.
  • Source:

Apple Developing Enterprise AI Servers Powered by M8 Ultra

  • What Happened: Reports surfaced that Apple is actively designing enterprise-grade AI servers packed with future M8 Ultra chips, targeting a potential 2029 release. If realized, this would mark Apple’s first major return to the server hardware market since the discontinuation of the Xserve in 2011.
  • Why It Matters: Signals Apple’s long-term vertical integration strategy extending from consumer edge devices directly into private-cloud enterprise AI compute infrastructure.
  • Source: Ars Technica Coverage

Research Papers

Continual Learning Mechanisms Compose for Long-Horizon Memorization — AI Research Community

  • Motivation: Addressing the degradation of long-horizon memory and catastrophic forgetting in multi-agent workflows and extended LLM task execution.
  • Key Innovation: Proposed compositional continual learning mechanisms that dynamically bind modular memory modules without destabilizing core parametric weights during extended agent runs.
  • Results: Demonstrated sustained zero-forgetting performance across multi-day conversational tasks and large code-base refactoring pipelines.
  • Paper: Hugging Face Papers |

Other Highlights

AgentGit: Version Control for Multi-Agent Sessions

  • Overview: Einsia introduced AgentGit, an open-source platform designed to save, resume, fork, and hand off active AI agent sessions. As multi-hour agent workflows become standard in engineering teams, AgentGit enables seamless context collaboration and session reproducibility across developers.
  • Link:
Share on:
Featured Partners

© 2026 Communeify. All rights reserved.