OpenAI will automatically watermark ChatGPT outputs in the European Union, driven by regulatory requirements.
What shipped
57 launch and release stories, grouped by the team behind them. Detection runs on published headlines and summaries — nothing is inferred beyond what was reported.
OpenAI
The AI research organization publishes a batch of 722 mathematical manuscripts, covering 372 result families and solving hundreds of open questions.
The new tool will help teenagers organize college applications and provide guidance throughout the process.
OpenAI is launching a new ad format for ChatGPT, featuring images of sponsored products and services.
Three months ago Dwarkesh, who has been posting incredible blogs and episodes about RL, posted a framing question for his video essay on RLVR which upset a…
Top NewsAnthropic and OpenAI race to release smarter and cheaper modelsSources:Anthropic launches Claude Opus 5.5 with stricter safeguards for…
Today is the 20 year anniversary of Sam Altman’s first startup, and fittingly OpenAI the consumer AI company is so back (as is OpenAI the AI Cloud and…
AI ChatGPT co-creator launches a new kind of AI PLUS: How to add models to Codex, Claude Code Good morning, AI enthusiasts, and welcome to our 4,056 new readers. Diogo Almeida helped build the methods that taught AI to talk to people — the research behind ChatGPT.
Top NewsGPT-6 Astra Is Here—and OpenAI Thinks It May Kick Off the AGI EraRelated:OpenAI begins rolling out Astra model after warning of its advanced…
Don't pick one. Run GLM-5.3 first, escalate to GPT-5.6 Sol when the tests fail. That cascade solves 85.9% of DeepSWE tasks at \$6.61 each. Sol alone solves 72.7% at \$8.37.
Today, we are excited to announce the release of Qwen3 , the latest addition to the Qwen family of large language models. Our flagship model, Qwen3-235B-A22B , achieves competitive results in benchmark evaluations of coding, math, general capabilities, etc., when compared to other top-tier models such as DeepSeek-R1, o1, o3-mini, Grok-3, and Gemini-2.5-Pro.
ComfyUI
The latest release of ComfyUI, a GitHub project, brings several updates and improvements, including support for LynnReal light Minimax-H3 vae and updated embedded docs.
MiniMax H3 now runs natively in ComfyUI 0.30.0 or later. The official templates cover text-to-video, image-to-video and multimodal reference-to-video, producing video and synchronized 32 kHz stereo audio in one MP4.
AlphaSignal
Subtopic Mixture Of Experts · Long Context Takeaways − Kolibri-1 is a 78B MoE with 3.46B active parameters, Apache 2.0, German and English focus. Context window validated up to 1,048,576 tokens, native 262,144, no position scaling tricks required.
NVIDIA
Takeaways − Red Hat AI released an FP8 quantized build of NVIDIA Nemotron 3.5 Lightning 30B A3B. Cuts GPU memory and disk by roughly 50% versus the BF16 reference weights.
Fine-Tuning NVIDIA Nemotron for Saudi Arabic Dialects, with a Path to Other Languages | NVIDIA Technical Blog Fine-Tuning NVIDIA Nemotron for Saudi Arabic Dialects, with a Path to Other Languages By Imane Khaouja , Amine El Khair , Meshari Alaeena , Zahra Al-Kaf and Abdulrahman Alkhamees NVIDIA Nemotron 3.5 ASR supports multilingual streaming transcription across 40 language-locales, but deployment-specific dialects...
Lower the Cost of Building and Running Visual AI Agents with NVIDIA VSS Blueprint 3.3 | NVIDIA Technical Blog Lower the Cost of Building and Running Visual AI Agents with NVIDIA VSS Blueprint 3.3 By Hassan Moustafa , Debraj Sinha and Ashwani Agarwal The NVIDIA Metropolis Blueprint for Video Search and Summarization (VSS) 3.3 connects vision-language models such as NVIDIA...
Add Runtime Controls to AI Agents with NVIDIA OpenShell | NVIDIA Technical Blog Add Runtime Controls to AI Agents with NVIDIA OpenShell NVIDIA OpenShell 0.1.0 provides an open-source runtime that enforces which systems and data an AI agent can access without rewriting the agent.
A warped Manhattan is waiting in the cloud this week. Remedy Entertainment’s CONTROL Resonant brings Dylan Faden’s extraordinary abilities and a paranatural…
NVIDIA Nemotron 3.5 Lightning is now available on Ollama, and it runs completely on your own device. It's a 30 billion parameter (3B active) open model from NVIDIA built for agents that stay running: gathering context, calling tools, and working through multi-step tasks.
Unveiling what it describes as the most capable model series yet for professional knowledge work, OpenAI launched GPT-5.2 in December.
Cohere
Subtopic Dpo · Fine Tuning · Distillation Takeaways − Cohere released North Small Translate , a 218B / 25B-active MoE translation model with open weights. Scores 83.60 on WMT26 across 50+ languages, beating DeepL NextGen (81.37) and Google Translate (68.20).
State-of-the-art enterprise retrieval: Embed 5 Pro achieves the highest average score of any model we tested, particularly across financial datasets, parsed PDFs, and visually rich documents. A new Fast tier: Embed 5 Fast brings strong retrieval quality to latency, and cost-sensitive workloads, at $0.08 per million tokens.
DeepSeek
Our 258th episode with a summary and discussion of last week’s big AI news!Recorded on 09/26/2026 ; as usual, I am sorry this is coming out late, next…
Meta
Meta now lets you make your own Muse gadgets that feature the company's new AI agent with code that the company open sourced.
Meta founder and CEO Mark Zuckerberg shared the following: We believe superintelligence will create significant new opportunities for all people and…
Muse Glimmer from Meta Superintelligence Labs is now available Muse Glimmer , Meta's newest open model and the first released by Meta Superintelligence Labs, is now available on Ollama . It's a 30B multimodal model purpose-built for agent workloads that run locally with a 128K+ context length, released under the Apache 2.0 license.
Improved performance and model support with GGUF Ollama 0.30 is now available with improved performance and GGUF model compatibility through llama.cpp . This augments Ollama's MLX engine on Apple silicon, bringing support to more models on a wider range of hardware.
Connect 2024: The responsible approach we’re taking to generative AI Connect 2024: The responsible approach we’re taking to generative AI Today at Connect 2024, we shared updates for Meta AI features and released Llama 3.2, a collection of models that includes new vision capabilities as well as lightweight models that can fit on mobile devices.
Runway
Runway said it is testing Praxis-1 on a variety of embodiments and environments to identify and close potential gaps before moving to general availability. | Source: Runway Runway AI Inc.
The latest AI news we announced in September 2026 Here’s a recap of some of our biggest AI updates from September, including Gemini 4 Argon, new Connected Apps in the Gemini App, and WeatherNext 3. Your browser does not support the audio element.
GDM last shipped a larger-than-Flash model in February (3.1 Pro), and after successive incremental 3.x Flash versions and the big GDM management shakeup last…
Intro Weaviate v1.39.3 includes a fix for a high severity credential disclosure vulnerability in Weaviate's Google-backed modules, where an unvalidated…
Google promised Gemini 3.5 Pro in June, but it spent the summer trotting out smaller Flash models.
Google.org announced on Sept. 15 that the MIT Transit Lab is a recipient of $2.1 million in funding — one of only 15 projects selected in the worldwide…
Introducing Gemini 3.8 Live with Live Avatar Gemini 3.8 Live with Live Avatar brings real-time visual presence to Gemini’s conversational AI. By natively coupling our live dialogue capabilities with low-latency streaming video, Live Avatar enables a more natural and intuitive conversational experience for enterprises and their users.
Gemini 3.8 Flash TTS and Gemini 3.8 Flash-Lite TTS are our most expressive audio generation models yet. Generate custom character voices and direct scene dialogue across Google AI Studio, Gemini API, Gemini Enterprise, Gemini Notebook, and Google Vids.
Introducing Gemini 3.8 Live and 3.8 Live Extended Thinking Gemini 3.8 Live and Gemini 3.8 Live Extended Thinking are our most advanced live dialogue models yet. Major upgrades in intelligence and parallel reasoning make them more intuitive to collaborate with and use to execute complex tasks using your voice.
SPONSORED BY ODSC AIODSC AI West 2026 runs October 27–29 in San Francisco and virtually, with 300+ sessions covering agentic AI for enterprise, personal…
Introducing Gemini 3.8 Flash and 3.8 Flash Cyber Our newest Gemini models deliver next-generation intelligence for agentic workflows and cybersecurity.
Our 255th episode with a summary and discussion of last week’s big AI news!Recorded on 08/26/2026Hosted by Andrey Kurenkov and Jeremie HarrisFeel free…
Note from Andrey: apologies for the newsletter not having resumed yet - the next release is going to come tomorrow, and it will resume weekly cadence…
Perplexity
Fetch the complete documentation index at: /llms.txt Use this file to discover all available pages before exploring further. The Decisions API is billed at $0.04 per million input tokens.
Anthropic
GLM-5.3 and the spread of advanced cyber capabilities Cole McFaul, Robert Xiao, Tripp Gallagher Five months ago, we announced Claude Mythos Preview, the first AI model that could autonomously build sophisticated, end-to-end cyber exploits.
Introducing Claude Sonnet 5.5, the second model in the Claude 5.5 family. It’s a clear upgrade over Claude Sonnet 5, runs 30%+ faster, and costs up to 30% less for most work.
Author(s): Chew Loong Nian – AI ENGINEER Originally published on Towards AI.
Hi folks, Keshav here. Ben’s travelling today, so you’re stuck with me.We have three new models to talk about:Claude Opus 5.5 - it’s the…
We’re introducing Claude Opus 5.5, the first model in our new Claude 5.5 family. It performs at the level of Claude Fable 5.1 on most work and costs 40% less to run than Opus 5.
GLM-5.3 and Claude Fable 5 finish within noise of each other on DeepSWE accuracy, but GLM-5.3 costs a fifth as much per task and wins every multi-attempt metric. When two models are this close on quality, the price gap becomes the entire decision.
Kling AI
Warning : Undefined variable $stocks in /var/www/briefs.co/htdocs/wp-content/plugins/oxygen/component-framework/components/classes/code-block.class.php(133) : eval()'d code on line 448 Warning : foreach() argument must be of type array|object, null given in /var/www/briefs.co/htdocs/wp-content/plugins/oxygen/component-framework/components/classes/code-block.class.php(133) : eval()'d code on line 448 Warning : Undefined variable $funds in /var/www/briefs.co/htdocs/wp-content/plugins/oxygen/component-framework/components/classes/code-block.class.php(133) : eval()'d code on line 472 Warning : foreach() argument must be of type array|object, null given in /var/www/briefs.co/htdocs/wp-content/plugins/oxygen/component-framework/components/classes/code-block.class.php(133) : eval()'d...
Alibaba
Alibaba s Qwen team shipped Qwen-Image-2.1 on September 20, 2026, and it changes the calculus for anyone running AI image generation on their own hardware in Australia. The 20-billion-parameter model handles text-to-image generation, multi-reference editing with up to 10 source images, and native transparent PNG output from a single checkpoint.
Last Updated on September 25, 2026 by Editorial Team Author(s): Ankit Agrawal Originally published on Towards AI.
Together AI
A global fintech runs its coding assistant on GLM 5.2 through Together's Dedicated Model Inference, handling spiky, engineering-hours traffic that static capacity planning couldn't keep up with. With DMI, the customer's engineers scale endpoints, roll out models, and test changes themselves, no tickets, no waiting on Together.
Weaviate
Weaviate v1.39 is now available open-source and on Weaviate Cloud. Two search features reach general availability in this release: the Boost API for…
Stability AI
Stable Diffusion 3.5 is the model Stability AI still points to as its default getting started release for local installs, even after the company announced Stable Diffusion 4 in two tiers — Base and Ultra — on April 6, 2026, priced through its API at $0.01 per credit.
Meet Stable Audio 3.0, the model family built for artistic experimentation with open-weight models We're releasing Stable Audio 3.0 , a model family with open-weights music models that are trained on fully licensed data.
xAI
Note from Last Week in AI (Andrey): I’m back! And i’m sorry for putting the substack on a silent pause, work got a bit too overwhelming so I fell…