Everything AI Released in October 2026

27 releases covered live on the show, led by OpenAI Dots, GPT-6.1 Sol — every model, product, paper and tool that mattered, with links and our analysis.

What AI products launched in October 2026?

27 AI releases shipped in October 2026, 27 of them in the latest week (Oct 1, 2026), led by OpenAI Dots, GPT-6.1 Sol, Claude Sonnet 5.5, Gemini 4 Argon, CoreWeave Serverless GPUs. ThursdAI — the weekly AI news podcast hosted by Alex Volkov — tracked every entry below with source links, key numbers and episode coverage (all covered live on the show).

What was the biggest AI story in October 2026?

CoreWeave launches serverless GPUs: GPU sandboxes by the hour, no contract. Announced live on ThursdAI by Deok Filho: serverless GPUs on CoreWeave, GPU sandboxes with untrusted code execution. ThursdAI covered it live on the show, with primary sources linked below.

What open-source AI models were released in October 2026?

1 open-weights model shipped in October 2026, led by Holo4. Each entry below links the weights and the episode segment where we covered it.

Which AI companies shipped in October 2026?

16 companies shipped AI releases in October 2026; the most active were CoreWeave, OpenAI, AMD, Anthropic, Cloudflare, ElevenLabs. Every launch below has primary-source links and ThursdAI's live episode analysis.

October 2026 verdict table: the launches that mattered and who they're for
ReleaseBest forWhy it mattersKey number
Serverless GPUs · CoreWeave Infra & GPU planners CoreWeave launches serverless GPUs: GPU sandboxes by the hour, no contract Per hour billing, no contract or commitment
Claude Sonnet 5.5 · Anthropic Developers & coding agents Anthropic ships Claude Sonnet 5.5, beating Opus 5.5 on Terminal-Bench 4.0 70.6 Terminal-Bench 4.0 (Opus 5.5: 66.4)
GPT-6.1 Sol · OpenAI Developers & coding agents OpenAI releases GPT-6.1 Sol: near-Astra intelligence, cached input 95% off $2 / $10 per 1M input / output tokens
AMD–World Labs acquisition Robotics tinkerers AMD to acquire Fei-Fei Li's World Labs for about $8.2B in stock ~$8.2B in AMD stock
Gemini 4 Argon · Google DeepMind Model evaluators Google's Gemini 4 Argon tops Text Arena and the Vals Index, trusted testers only #1 Text Arena
ElevenLabs v4 Voice-agent builders ElevenLabs releases v4 and v4 Turbo speech models 90 languages
UltraFast mode · OpenAI Developers & coding agents OpenAI launches UltraFast mode on Cerebras with a $500 Pro tier 8x faster in Codex (up to)
Span-01 · Respan Agent builders Respan's Span-01 decision model claims half the price of Jev 2x cheaper than Jev (claimed)
Forge · CoreWeave Agent builders CoreWeave launches Forge, the whole agentic loop in one place $60 per month for Pro
NVIDIA Vera CPU · CoreWeave Agent builders CoreWeave to offer NVIDIA Vera, the first CPU built for AI agents 20,000+ agents per Vera CPU rack
Holo4 · H Company Agent builders H Company releases Holo4 computer-use models —
D1 · Liquid AI Agent builders Liquid AI releases D1, its first decision model —

🧠 New Models 8

Anthropic
New Models

Claude Sonnet 5.5

Anthropic ships Claude Sonnet 5.5, beating Opus 5.5 on Terminal-Bench 4.0

A week after Opus 5.5, Claude Sonnet 5.5 is over 30% faster than Sonnet 5 and up to 30% cheaper for most work at $2 / $10 per million tokens. It scores 70.6 on Terminal-Bench 4.0, above Opus 5.5's 66.4, and is available on Claude's free tier.

70.6 Terminal-Bench 4.0 (Opus 5.5: 66.4)30%+ faster than Sonnet 5$2 / $10 per 1M input / output tokens
OpenAI
New Models

GPT-6.1 Sol

OpenAI releases GPT-6.1 Sol: near-Astra intelligence, cached input 95% off

Five days after GPT-6 Sol, GPT-6.1 Sol keeps the $2 input / $10 output price but cuts cached input 95% to $0.10 per million tokens. Artificial Analysis puts it 1 point below GPT-6 Astra at about 72 cents a task versus over $3, it beats Astra on DeepSWE at roughly a sixth of the cost, and it is the new default in Codex.

$2 / $10 per 1M input / output tokens$0.10 per 1M cached input tokens (95% off)$0.72 per task on Artificial Analysis vs over $3 for Astra

🚀 Products & Apps 8

CoreWeave
Products & Apps

Forge

CoreWeave launches Forge, the whole agentic loop in one place

Forge combines run, observe, curate, improve and evaluate in one product, with Weights & Biases (W&B Models included), OpenPipe's post-training, marimo notebooks and CoreWeave Sandboxes. It has a free tier, and Pro starts at $60 a month with a 30-day trial.

$60 per month for Pro30 days free Pro trial
CoreWeave
Products & Apps

NVIDIA Vera CPU

CoreWeave to offer NVIDIA Vera, the first CPU built for AI agents

CoreWeave is bringing NVIDIA's Vera CPU to its cloud for agent workloads. Deok Filho's new metric is agent packing: one rack of Vera CPUs runs over 20,000 agents at once on about 11,000 cores.

20,000+ agents per Vera CPU rack~11,000 cores per rack
CoreWeave
Products & Apps

Serverless GPUs

CoreWeave launches serverless GPUs: GPU sandboxes by the hour, no contract

Announced live on ThursdAI by Deok Filho: serverless GPUs on CoreWeave, GPU sandboxes with untrusted code execution. Sign up at forge.coreweave.com, ask to have it enabled, and pay per hour with no contract and no commitment. Private preview started September 30, with more GPU SKUs planned.

Per hour billing, no contract or commitmentSep 30 private preview start
OpenAI
Products & Apps

Dots

OpenAI launches dots, always-on ChatGPT agents on GPT-6 Astra

OpenAI's DevDay headline: a dot is an always-on agent in ChatGPT, powered by GPT-6 Astra, with its own computer and browser in the cloud and access to the 4,000+ apps built for ChatGPT. It works 24/7 and can be reached from Slack, Teams or a phone-call style interface. Pro only for now, including the $100 plan.

4,000+ ChatGPT apps a dot can use

✨ Major Features & Updates 5

CoreWeave
Major Features & Updates

Serverless distillation

CoreWeave adds serverless model distillation to Forge

Forge's serverless side covers inference, RL and SFT, and now model distillation: train a small student on a frontier teacher for your task without managing Kubernetes or a cluster.

OpenAI
Major Features & Updates

UltraFast mode

OpenAI launches UltraFast mode on Cerebras with a $500 Pro tier

A new $500 Pro tier adds UltraFast mode, OpenAI's models served on Cerebras: up to 8x faster in Codex and around 300 tokens a second for GPT-6 Astra, at 6x the price. The $200 Pro plan returns with half the usage.

8x faster in Codex (up to)~300 tokens per second for GPT-6 Astra$500 per month Pro tier

🔌 APIs & Platforms 1

OpenAI
APIs & Platforms

Decisions API

OpenAI previews a Decisions API for fast fixed-choice answers

The Decisions API takes a question and a fixed set of answers and returns one answer, fast. It runs on Luna, accepts multimodal input and is waitlisted, and it closely resembles TypeSafe's Jev decision model.

🛠️ Dev Tools 3

📦 Datasets 1

Nisten Tahiraj
DatasetsOpen weights

Opus 5.5 synthetic medical dialogue dataset

Nisten's synthetic doctor-patient dataset hits #2 on Hugging Face

Nisten released a 70MB synthetic dataset generated with Claude Opus 5.5 by an agent loop running 100 agents at a time across 2,200 diseases: doctor-patient role-plays where the patient hides something and the doctor has to find it, filtered three times. It reached #2 in Hugging Face datasets and #6 overall.

#2 Hugging Face datasets2,200 diseases covered

🤝 Acquisitions 1