← Back to feed

Models & Releases

L4 · DeveloperModels & ReleasesHacker News· 8/27/2026
Gemini-3.5-Transcribe

Google launches Gemini 3.5 Transcribe, a speech-to-text model with 4.0% WER for streaming and 2.6% for non-streaming use cases, now available via API.

L3 · BuilderModels & ReleasesHacker News· 8/27/2026
Gemini Omni 1.1 Flash

Google released Gemini Omni 1.1 Flash, a production-ready video generation model with new controls for scene extension, frame interpolation, 4K upscaling, and faster prototyping.

L3 · BuilderModels & Releases@GoogleAI· 8/27/2026
Google AI (@GoogleAI): Meet Gemini Omni 1.1 Flash ⚡️ Our newest multimodal model for video generation and editing. It now features your favorite creative controls from Veo, plus brand new capabilities. Enjoy features like

Google released Gemini Omni 1.1 Flash with 10-second scene extension capability, 4K upscaling, and improved video generation controls.

L3 · BuilderModels & ReleasesSimon Willison's Blog· 8/26/2026
Qwen3.8-Flash-Next

Qwen3.8-Flash-Next is a new 125B token multimodal Mixture-of-Experts model that serves as an early preview of the Qwen4 architecture, with only 6B active parameters for performance efficiency.

L3 · BuilderModels & ReleasesTechCrunch AI· 8/26/2026
Surprise: Z.ai is the AI lab behind the mysterious Ox Alpha model

Z.ai revealed it created the mysterious Ox Alpha model, an open-weight reasoning model for coding and agentic work that tops benchmarks.

L3 · BuilderModels & ReleasesTechCrunch AI· 8/26/2026
Ex-Meta scientists want to bring visual AI to the factory floor

Perceptron AI launched Isaac 0.5, an open-weight vision model that helps robots perceive, reason, and act in industrial settings like warehouses and factory floors.

L4 · DeveloperModels & Releases@GoogleAI· 8/25/2026
Google AI (@GoogleAI): https://t.co/LJGIbM3Ksy

Google's WeatherNext Cyclones model, used in real-time by the U.S. National Hurricane Center, provided 5-day advance warning for a Category 5 hurricane and offers accuracy improvements equivalent to a decade of meteorological progress.

L5 · ResearcherModels & ReleasesHugging Face Blog· 8/25/2026
Granite 4.2 LLMs: How They're Built

IBM releases Granite 4.2 reasoning LLMs in 3B, 8B, and 30B sizes with 512K context, agentic RL training, and Apache 2.0 license.

L4 · DeveloperModels & ReleasesTheSequence· 8/26/2026
The Sequence Learning Loop - Issue #921: Learn About DeepSeek New Model, the Env Harness Paper and the Amazing Etched

DeepSeek added vision capabilities to its V4 model, Google introduced the EnvHarness training framework, and Etched shipped its first AI inference hardware to Jane Street.

L4 · DeveloperModels & ReleasesHacker News· 8/26/2026
Z.ai confirms Ox Alpha is a new GLM-series model and will release its weights

Z.ai confirms Ox Alpha as a new GLM-series model and plans to open-source its weights.

L4 · DeveloperModels & ReleasesAI Business· 8/26/2026
Qwen 3.8 Flash-Next is Cheap, But There Are Complicating Factors

Alibaba's Qwen 3.8 Flash-Next offers low-cost inference, but enterprises must evaluate other performance metrics before adoption.

L1 · CuriousModels & Releases@sama· 8/26/2026
Sam Altman (@sama): i think we should do another party for our next model release, the 5.5 party was a lot of fun. what would make the next one awesome?

Sam Altman suggests throwing another party for OpenAI's next model release, referencing the fun of their GPT-5.5 celebration.

L1 · CuriousModels & Releases@sama· 8/25/2026
Sam Altman (@sama): we made a chip and it is fast

Sam Altman announces OpenAI has developed a custom AI chip described as fast, signaling potential hardware acceleration for model inference.

L2 · PractitionerModels & Releases@GoogleAI· 8/26/2026
Google AI (@GoogleAI): Today we’re introducing Gemini 3.5 Transcribe, our latest transcription model built for incredibly precise, smart dictation across your favorite apps and devices. Remember when traditional speech-to-

Google launches Gemini 3.5 Transcribe, a multimodal transcription model that filters filler words and formats speech across 85+ languages.

L3 · BuilderModels & ReleasesAI Business· 8/24/2026
Thomson Reuters’ New Model Could Inspire Other SaaS Vendors

Thomson Reuters developed a new AI model using an open-weight foundation model as its base architecture.

L3 · BuilderModels & Releases@GoogleDeepMind· 8/26/2026
Google DeepMind (@GoogleDeepMind): Gemini 3.5 Transcribe is our latest speech-to-text model for precise and intelligent transcriptions. 🧵 https://t.co/24IDhwBtl7

Google DeepMind released Gemini 3.5 Transcribe, a speech-to-text model with improved accuracy for phone numbers, custom vocabulary, and noise handling.

L2 · PractitionerModels & ReleasesTechCrunch AI· 8/23/2026
Who’s behind the new ‘stealth model’ Ox Alpha?

A mysterious 'stealth model' called Ox Alpha has been released on OpenRouter, sparking speculation about whether it's from Chinese developers or an unreleased Microsoft model.

L4 · DeveloperModels & Releases@ClementDelangue· 8/25/2026
clem 🤗 (@ClementDelangue): Who's excited? https://t.co/jiKFk0vwJn

Hugging Face announces Qwen3.8-Flash-Next, a preview of Qwen4 architecture, scheduled for release on August 26, 2026.

L4 · DeveloperModels & ReleasesOpenAI Blog· 8/24/2026
Advancing price-performance for developers with GPT‑5.6 in Kiro

OpenAI released GPT-5.6 in Kiro with improved price-performance for software development tasks like planning, building, and testing.

L4 · DeveloperModels & ReleasesHacker News· 8/23/2026
GLM-5.3 (open-weight) beat Anthropic/OpenAI models – for 1/5 the cost

An open-weight model, GLM-5.3, reportedly outperformed proprietary models like GPT-5.5 and Claude Sonnet/Opus in a leaderboard of real-world tasks while costing 80% less.

L4 · DeveloperModels & ReleasesHugging Face Blog· 8/20/2026
Up to 3.2x Faster Inference with LFM2.5-DSpark

LiquidAI releases DSpark draft models for LFM2.5 family, achieving up to 3.2x faster inference through speculative decoding while maintaining output quality.

L4 · DeveloperModels & ReleasesTechCrunch AI· 8/22/2026
Inherent, founded by DeepMind alumni, says its AI ‘teammate’ just outperformed Anthropic and OpenAI at replicating research

Inherent's Faraday AI agent outperformed Claude Opus 4.8 and GPT-5.5 at scientific paper replication using a smaller 27B parameter Qwen model.

L4 · DeveloperModels & ReleasesAI Business· 8/20/2026
Nvidia’s SONIC Teaches Humanoids to Move

Nvidia released SONIC, a model that enables humanoid robots to learn motion in real-time from human demonstrations.

L4 · DeveloperModels & ReleasesHacker News· 8/22/2026
GPT 5.6 Sol 20% price reduction

OpenAI reduced GPT-5.6 Sol pricing by 20% and introduced new features like reasoning_effort controls and realtime prompt caching.

L4 · DeveloperModels & ReleasesHacker News· 8/21/2026
DeepSeek-v4-flash-vision-exp

DeepSeek released v4-flash-vision-exp, a multimodal model that accepts images alongside text with three integration methods: base64 encoding, external URLs, and Files API references.

L4 · DeveloperModels & ReleasesHugging Face Blog· 8/19/2026
LFM2.5 Q4\_0 Checkpoints from Quantization-Aware Distillation

LiquidAI releases Q4_0 quantized checkpoints for LFM2.5 models using quantization-aware distillation, recovering 97% of BF16 accuracy while maintaining 4-bit efficiency.

L4 · DeveloperModels & ReleasesHacker News· 8/17/2026
GPT-5.6 Sol Pricing Cut by 50%

OpenAI cuts GPT-5.6 Sol API pricing by 50% to $2.50/$15 per 1M tokens, making it more accessible for developers building complex reasoning and coding applications.

L5 · ResearcherModels & ReleasesarXiv· 8/13/2026
AlayaWorld: Interactive Long-Horizon World Modeling - Full Technical Report (v1.1)

AlayaWorld v1.1 introduces major architectural revisions for interactive long-horizon world modeling, including a new 3D point-cache renderer and redesigned conditioning pipeline.

L4 · DeveloperModels & ReleasesSimon Willison's Blog· 8/16/2026
Qwen 3.8 27B is excellent, but it defaults to wildly overthinking things

Qwen 3.8 27B is a new open-source vision-capable LLM that defaults to an 'xhigh' reasoning effort setting, causing extensive overthinking even for simple tasks.

L4 · DeveloperModels & ReleasesTheSequence· 8/19/2026
The Sequence Frontier Learning - Issue 917: Understanding DeepSeek V4-Pro, GLM-5.3, NVIDIA Nemotron 3.5 Lightning and NeMo Switchyard

Analyzes DeepSeek V4-Pro, GLM-5.3, NVIDIA Nemotron 3.5 Lightning, and NeMo Switchyard in technical depth, focusing on their architectural innovations and benchmark performance.