AI news and announcements

Recent AI developments, summarized by dstilled and linked to their original announcements.

How NVIDIA GPUs Help Accelerate OpenAI’s GPT-6 Astra Ultrafast

The new mode delivers up to 8x faster token generation than Astra Standard mode for API, ChatGPT Work, and Codex users.

Cursor adds GLM 5.3 and GLM 5.3 Flash

GLM 5.3 Max ranks as the highest-scoring open-weight model on the CursorBench 4.0 benchmark.

Grok Build launches Agent Dashboard for managing multiple agents

The dashboard offers a centralized interface to dispatch agents, execute tasks in parallel, and isolate agents in dedicated git worktrees.

Anthropic adds mod support to Claude Code

Developers can now customize Claude Code's interface, behavior, and features using TypeScript plugins installed directly via the CLI or desktop app.

Grok Bot Adds Proactive Assistance Suggestions

The primary bot will proactively identify tasks and offer to handle them without counting suggestions against usage quotas.

Our first streaming transcription model debuts at no. 1 on Artificial Analysis | Microsoft AI

MAI-Transcribe-2-Streaming provides real-time transcription across 60 languages with 100ms partial hypothesis latency, priced at $0.54 per audio hour through year-end.

Perplexity Adds Interactive Charts and Visualizations to Threads

The feature integrates TradingView Lightweight Charts to display financial data including candlesticks, volume, and moving averages directly inside threads.

ElevenLabs expands to the Netherlands with Amsterdam office

ElevenLabs established an office in the Netherlands and plans to triple its local GTM and engineering headcount this year.

Gemini 4 Argon: our next era of frontier intelligence

The model features a 1M-token output limit and will launch at $2 per million input and $10 per million output tokens.

Perplexity releases pplx-embed-v2-context-9b-preview contextual embedding model

The model trains contextual chunk retrieval by distilling relevance scores from a query-aware context compression model instead of single gold chunks.

Grok Bot Adds Cursor Handoff and GitHub PR Management

Bots can now delegate coding tasks to Cursor, manage pull requests using GitHub and Origin plugins, and generate video demonstrations of builds.

OpenAI Adds Shareable Profiles to ChatGPT

Google Labs launches Skills in Gemini and sunsets Opal

Skills lets Gemini users automate repetitive tasks and save custom instructions directly inside chat.

Disrupting a coordinated model-distillation campaign

OpenAI attributed a core cluster of the extraction attempts involving over 15,000 accounts to individuals associated with Moonshot AI, the developer of Kimi.

Google launches stackable skills in Gemini to replace Gems

Skills let users save reusable instructions with reference files, automatically stack them for complex tasks, and will fully replace Gems starting in November.

ElevenLabs reaches $22bn valuation after employee tender

Wellington and T. Rowe Price led the transaction, doubling ElevenLabs' valuation since its Series D funding round in February.

How Open Science Can Help Researchers Prepare for the Next Pandemic

The predicted structures were generated using AlphaFold2 optimized with NVIDIA BioNeMo Inference Runtime and added to the AlphaFold Database.

Introducing SynthID Bio

The watermarking tool embeds verification signals directly into AI-designed sequences to help DNA synthesis providers screen orders and track biological provenance.

ElevenLabs launches in Belgium and opens office in Brussels

Falke Van Onacker will lead the new Brussels office as ElevenLabs plans to triple its local workforce this year.

Cursor Adds Inline Charts and Diagrams via /visualize Command

The feature is available in the Agents Window to analyze data and render charts directly in chat.

OpenAI releases Decisions API powered by GPT-6 Luna

The API allows applications to classify content, route requests, or determine agent actions using text or image inputs.

OpenAI updates Codex CLI with voice and parallel agent workflows

The update adds bidirectional voice interaction, session forking into managed worktrees, and tools to monitor parallel agent tasks inside the terminal.

OpenAI Launches Codex Cloud Environments for Persistent Tasks

The feature is rolling out to ChatGPT Plus, Pro, Business, and Enterprise subscribers.

OpenAI releases Ultrafast mode for Astra

The mode runs up to 8x faster than Astra Standard and 4x faster than Astra Fast in Codex.

Introducing dots

The GPT-6 Astra-powered agents run on dedicated cloud computers, connect with over 4,000 apps, and are rolling out to Pro, Business Premium, and Enterprise subscribers.

Introducing GPT-6.1 Sol

The model approaches GPT-6 Astra benchmark performance across coding and professional tasks at $2 per million input and $10 per million output tokens.

Sign in with ChatGPT | Devin

Subscribers to eligible ChatGPT Plus or Pro plans can connect their accounts to cover OpenAI model usage across Devin Cloud, Desktop, and CLI.

OpenAI releases GPT-6.1 Sol for agentic coding and computer use

The model costs $2 per million input tokens, $10 per million output tokens, and $0.10 per million cached input tokens.

GPT-6.1 Sol is now available in Devin | Devin

The model is live in Devin Desktop and Devin CLI, matching previous benchmark performance while cutting per-task costs by up to 81%.

OpenAI launches dots, an always-on developer agent

Powered by GPT-6 Astra, the agent runs in an isolated cloud environment to triage bugs, fix failing builds, and submit pull requests.

ElevenLabs launches ElevenAgents on OpenAI Marketplace

OpenAI enterprise customers can purchase and deploy conversational agents using their existing OpenAI financial commitments.

Cognition and MongoDB partner to modernize enterprise infrastructure in months, not years | Devin

The integration brings Cognition's Devin AI agent into MongoDB's Application Modernization Platform to automate rewriting legacy code, queries, and data access layers.

Perplexity Introduces Automations in Perplexity Computer

Automations perform scheduled or event-triggered actions integrated with memory, skills, and workplace apps.

ElevenLabs launches Eleven v4 Turbo in ElevenAgents

The model provides speech synthesis with a median inference latency of about 100 ms across more than 90 languages at 3.3 cents per minute through October 12th.

The Future Is for Everyone: Muse for Small Business

The personal AI agent adds integrations for Instagram analytics, Facebook Pages, Meta ad accounts, and third-party services like Canva.

ClaudeDevs Announces Release of Claude Sonnet 5.5

Anthropic published a developer guide covering migration from Sonnet 5, tuning effort, and usage in Claude Code.

Cursor Adds Sonnet 5.5 Support

xAI launches Team Bots in public beta for Enterprise

Team Bots are available in public beta for Teams and Enterprise tiers.

Team Bots: AI coworkers that learn from your team

The feature lets Teams and Enterprise users deploy shared Grok Bots that integrate with workplace apps while keeping individual conversations private.

Anthropic releases Claude Sonnet 5.5

The model runs over 30% faster and reduces per-task costs by up to 30% through improved token efficiency compared to Sonnet 5.

Devin is now up to 40% more cost-efficient | Devin

Usage costs are reduced by 30-40% in Fusion and Normal modes, 15-20% in Ultra, and up to 70% in Devin Review.

The Lenfest Institute grows landmark program with expanded OpenAI support

The commitment doubles OpenAI's previous support to expand the Lenfest AI Collaborative and Fellowship Program across US newsrooms.

Mistral Opens Munich Hub to Advance Industrial AI in Germany

The hub houses dedicated research teams collaborating with BMW on crash simulations, Siemens Energy on industrial AI, and TUM on aerodynamics.

ElevenLabs releases Eleven v4 and Eleven v4 Turbo

The models support over 90 languages and inline performance tags, with the Turbo variant delivering approximately 100 ms median inference latency.

Eleven v4: Our most expressive text-to-speech AI model yet

The text-to-speech models support over 90 languages, inline emotion prompting, and a median time to first speech of ~150ms on the Turbo variant.

Launching Meta Enterprise Platform

Former MongoDB CEO CJ Desai joins Meta as Chief Enterprise Platform Officer to lead the new enterprise business offering tools like Muse API and Muse Code.

TypeSafe AI Reopens Jev Signups and Removes Free Credits

New signups are open following a capacity increase, but new accounts no longer receive free credits.

OpenAI fixes image understanding bug in GPT-6 Sol and Luna

The fix restores degraded visual task performance across the API and Codex, including computer use workflows.

Grok Bot launches Finance integration via Plaid

The read-only integration links bank, card, and investment accounts to assist users with spending and investments.

Claude launches developer portal to submit and manage plugins

Developers can submit single MCP connectors or GitHub plugin bundles directly at claude.ai/directory/manage/new and track review status.

Claude Code adds graceful wrap-up on session limits

The tool now draws a small allowance from the weekly limit to complete active edits instead of cutting off mid-task at the 5-hour session cap.

Anthropic Uses Claude to Solve Nine-Loop Physics Calculation

Claude ran largely unsupervised for days in Claude Science to calculate a nine-loop scattering amplitude at a total cost of a few thousand dollars.

ElevenLabs launches ElevenLabs for Students program

University students aged 18 and older in the US, Canada, EU27, Australia, and the UK receive free platform and API access.

FLUX 3 Action: A 7B World Action Model for Robot Control

The open-weight 7B World Action Model jointly predicts future video and robot actions, achieving up to a 42.2% success rate on RoboLab.

Perplexity launches Fast Search in Search API

The new tier runs on the custom Rust engine Photon, returning 95% of results within 230 ms and cutting agent task costs by 68%.

Anthropic resumes charging for safeguard-blocked requests in three categories

The policy applies to blocks in biology, distillation attacks, and frontier LLM development to counter coordinated system attacks.

Perplexity releases Portable Computer for AMD Ryzen AI Max processors

The feature is available in the Perplexity app for Windows to all Consumer and Enterprise Pro and Max subscribers.

Introducing Gemini 3.8 Live with Live Avatar

The feature combines real-time video generation and speech in Gemini Enterprise, providing synchronized lip-syncing and expressions across 97 languages.

Kimi releases Kimi Code Desktop 1.0.3

The update adds built-in browser previews for local HTML files, Composer slash command hints, and fixes an embedded server vulnerability.

ChatGPT Ads expands to Southeast Asia and Taiwan

The expansion covers Indonesia, Malaysia, the Philippines, Singapore, Thailand, Vietnam, and Taiwan, bringing total ChatGPT Ads availability to more than 60 countries.

New Features for Meta Ray-Ban Display

The smart glasses add voice-driven digital avatars for video calls, landmark-based navigation, Dolby Atmos capture, and online ordering across five new countries.

Introducing Meta VR Glasses: A Cinema, Courtside Seat, and Workspace in Just 100 Grams

The 100-gram glasses offload compute and battery to a tethered puck, feature 5K micro-OLED displays, and ship in Spring 2027.

Introducing Ray-Ban Meta Audio and More AI Glasses Styles

Ray-Ban Meta Audio glasses start at $349 with 12-hour battery life, while Ray-Ban Meta (Gen 3) starts at $449 with 3K video recording.

Anthropic launches Claude Code cloud sessions out of preview

The feature executes tasks on Anthropic-hosted infrastructure without requiring a local machine to remain open, offering one-time credits of $100 for Pro and $250 for Max users.

Cursor Launches Rollouts and Updates Security Reviewer

Both tools are available immediately on Teams and Enterprise plans, with free Rollouts credits provided for 10 days.

Introducing MentalHealthBench

The open benchmark uses expert-authored weighted rubrics to assess model responses across non-acute, high-acuity, and emergency conversations.

Anthropic launches Claude Marketplace for tools, agents, and partners

Users can access connectors, buy third-party agents, and connect with enterprise service partners directly through the directory.

Anthropic uses Claude to discover novel bacteriophage enzyme system

The newly identified system contains repeating DNA structures similar to CRISPR that may cut, copy, and paste DNA.

Devin now works across Microsoft Teams and Microsoft 365 | Devin

The release adds direct messaging in Teams and six Microsoft 365 MCP connectors covering Outlook, Calendar, OneDrive, SharePoint, To Do, and Entra ID.

Google expands Beam video collaboration system across six countries

The system ships as HP Dimension with Google Beam across the US, Canada, UK, France, Germany, and Japan through 18 channel partners.

Google Docs adds NotebookLM notebooks as context sources

Users can reference NotebookLM notebooks alongside emails, chats, and files directly within Google Docs via Workspace Intelligence.

Google Gemini adds MCP connections for 13 third-party apps

The new slate of connections is rolling out directly in the Gemini app today.

Cursor Cuts Agent Token Costs by 7%

The reduction was achieved without degrading agent quality through tighter prompts, selective tool loading, better caching, and compressed file reads.

Gemini 3.8 text-to-speech says hello

The two text-to-speech models offer prompt-driven voice design across 100+ languages, 30-second voice replication with verbal consent checks, and line-by-line performance direction.

Google AI Studio Releases Gemini 3.8 Flash TTS Models

The models are available immediately through the Gemini API and in Google AI Studio for audio generation.

Grok Bot Adds Google Workspace Integrations and Custom Network Routing

Grok Bot now connects natively to Google Slides, Sheets, and Docs, and can route internet traffic through a user's own network.

Perplexity introduces hint-guided self-distillation to reduce tool-call errors

The method trains models on real-world sessions by using corrective hints during training to align hint-free next-token predictions with hint-guided outputs.

OpenAI updates prompt caching for GPT-6 API

Shared prompt prefixes reused within a 30-minute window now qualify for input token discounts of up to 90%.

Perplexity adds GPT-6 Sol to Perplexity and Computer

GPT-6 Sol is now the default Light option in Computer's effort selector.

Introducing GPT-6 Sol and Luna

The models cut API prices by 50% compared to GPT-5.6, costing $2/$10 per million tokens for Sol and $0.10/$0.50 for Luna.

Anthropic Makes Claude Opus 5.5 Available Today

Anthropic released Claude Opus 5.5 alongside a showcase of early community explorations spanning interactive interfaces, simulations, and generative tools.

GPT-6 Sol and GPT-6 Luna are now available in Devin | Devin

Both models are live in Devin Desktop and Devin CLI, delivering comparable or higher benchmark scores at lower task costs.

OpenAI Launches GPT-6 Sol and Luna Models

Both models are available via API at 50% lower prices than GPT-5.6 and are rolling out in Codex and ChatGPT Work.

Cursor Adds Claude Opus 5.5

The model leads CursorBench at 57.8% and costs 40% less per task than Opus 5.

Perplexity adds Claude Opus 5.5 to Perplexity Computer

The model serves as the Standard effort level, scoring 0.610 at $4.13 per task on Perplexity's WANDR benchmark.

Anthropic releases Claude Opus 5.5 model

The model matches Claude Fable 5.1 performance while generating output 30% faster and costing 40% less than Opus 5.

ElevenLabs releases Scribe v2 Medical speech-to-text model

The model cuts clinical word error rates by 35% compared to base Scribe v2 and is available via API starting at $0.22 per hour.

Scribe v2 Medical is now available to everyone

The clinical speech recognition model is generally available on the batch Speech to Text API using the model identifier scribe_v2_medical.

ElevenLabs — Introducing Eleven Multilingual v1

The new model supports French, German, Hindi, Italian, Polish, Portuguese, and Spanish across all subscription tiers on the ElevenLabs beta platform.

Eleven Multilingual v2

The new model expands speech synthesis capabilities across 29 languages.

TypeSafeAI pauses signups for Jev due to high demand

Existing accounts will continue to function normally while the team works to restore open access.

NotebookLM Releases Interactive Learning Overviews to All Users

The feature, located under Reports, combines source summaries with studio artifacts into an interactive hub for studying and topic exploration.

Bringing Devin Cloud to your terminal | Devin

Developers can now create, steer, resume, and watch Devin Cloud VM sessions directly from their local terminal using `devin --cloud`.

Advisory Group on Mathematics and Artificial Intelligence

Hosted at the Institute for Advanced Study, the independent, unpaid group will advise OpenAI on research standards and the dissemination of math capabilities.

ElevenLabs Releases Studio 4.0 Video Editor

The platform integrates multimodal generation directly into a rebuilt multi-track timeline alongside Studio Agent, an AI co-editor that executes edits from text prompts.

Perplexity Adds Video Generation via MiniMax and ByteDance Models

The feature generates finished video alongside copy and creative in the same thread for Pro and Max subscribers.

Introducing Grok 4.7

The model is available immediately via API, Cursor, and Grok Build starting at $2 per million input tokens and $6 per million output tokens.

Gemini rolls out Live Chat to Ultra mobile users

The feature enables real-time voice conversations with notebooks across roughly 100 languages.

Kimi Launches Kimi Code Desktop for macOS and Windows

The standalone desktop application includes multi-agent parallel execution, an integrated browser and terminal, and support for third-party model providers.

TypeSafe AI opens Jev to the public

All users receive $5 in starter credit, equivalent to roughly 120 million tokens.

Open the live feed