Skip to main content
Daily Brief · Archive

Wednesday, September 30, 2026

Every story published that day — lead first, then grouped by section.

News

7
Illustration for: Deploying an HSTU Generative Recommender with NVIDIA Dynamo-
News

Deploying an HSTU Generative Recommender with NVIDIA Dynamo-Triton | NVIDIA Technical Blog Deploying an HSTU Generative Recommender with NVIDIA Dynamo-Triton By Junyi Qiu , Shijie Liu , J Wyman and Mudit Aggarwal Generative recommender systems reformulate personalization as sequence modeling over user behavior, and NVIDIA recsys-examples now provides an end-to-end HSTU inference workflow with Dynamo-Triton.

NVIDIA Developer Blog10 min
Illustration for: Tracing Agent Harness Behavior with NVIDIA NeMo Relay
News

Tracing Agent Harness Behavior with NVIDIA NeMo Relay | NVIDIA Technical Blog Tracing Agent Harness Behavior with NVIDIA NeMo Relay Learn how to use execution traces to understand agent behavior and determine whether harness changes improve task outcomes.

NVIDIA Developer Blog13 min
Illustration for: OpenAI connects the dots on always-on agents
News

AI OpenAI connects the dots on always-on agents PLUS: Get started with ChatGPT dot, OpenAI's new agent Good morning, AI enthusiasts, and welcome to the 5,480 new readers who joined us yesterday. The always-on AI agent category has gotten pretty crowded this summer.

The Rundown AI7 min

Models & Research

13
Illustration for: Runway Research | Introducing Praxis-1 - Runway
Models & Research

Runway app for iPhone Runway app for Android An open-weight world action model that turns Runway's video pretraining into control for real robots. “Pick up the tennis ball and put it in the box.” Today we're announcing Praxis-1 , our first open-weight world action model.

RunwayML Blog3 min
Illustration for: Gemini 4 Argon: our next era of frontier intelligence
Models & Research

Gemini 4 Argon: our next era of frontier intelligence Gemini 4 Argon delivers frontier performance in complex workflows across real-world software engineering, enterprise knowledge work like legal and finance, and cybersecurity defense. SVP, Google DeepMind and Chief AI Architect, Google Google’s new Gemini 4 Argon model brings advanced reasoning to complex, long-horizon professional tasks.

Google DeepMind7 min
Illustration for: Can we predict the jobs robots will do? - Anthropic
Models & Research

We present a robot exposure index based on how well robots can perform job tasks today. Robots, which we define as autonomous physical machines that sense and act, can perform three-quarters of physical tasks in the US, making up 34% of working hours, but mostly in limited settings.

Anthropic News53 min
Illustration for: On the Effectiveness-Fluency Trade-Off in LLM Conditioning:
Models & Research

On the Effectiveness-Fluency Trade-Off in LLM Conditioning: A Systematic Study - Apple Machine Learning Research research area Methods and Algorithms , research area Speech and Natural Language Processing conference EMNLP content type paper published September 2026 On the Effectiveness-Fluency Trade-Off in LLM Conditioning: A Systematic Study Authors Iuri Macocco†, Pau Rodríguez Lopez, Arno Blaas, Luca Zappella, Marco Baroni†*, Xavier Suau...

Apple Machine Learning2 min
Illustration for: SCLATE: A Substrate for Continual-Learning Agent Training an
Models & Research

SCLATE: A Substrate for Continual-Learning Agent Training and Evaluation - Apple Machine Learning Research research area Methods and Algorithms , research area Tools, Platforms, Frameworks content type paper published September 2026 SCLATE: A Substrate for Continual-Learning Agent Training and Evaluation Authors Youngmok Jung, Sirajul Salekin, Henry Tran, Javier Movellan, Zhao Huang, Manjot Bilkhu Continual-learning agents are systems of models, harnesses,...

Apple Machine Learning2 min

Products & Tools

7
Illustration for: Decisions API - Perplexity
Products & Tools

Fetch the complete documentation index at: /llms.txt Use this file to discover all available pages before exploring further. The Decisions API is billed at $0.04 per million input tokens.

Perplexity10 min

Guides

5
Illustration for: A practical guide to cost optimization with Lakebase Postgre
Guides

Lakebase is cost-efficient by design because its separated storage and compute architecture lets branching, read replicas, and high availability share one storage layer, while serverless autoscaling and scale-to-zero mean you pay only for the compute you actually use.

Databricks Blog12 min

Open Source

1

Policy & Ethics

2