Skip to main content
Tag

safety

33 stories

Illustration for: AI Is Making a Mess of Nurses’ Schedules. They Say It’s a Sa
News

Over the past four months, Amber Retzloff, a critical care nurse in Florida, requested to work 50 specific 12-hour shifts. But over half the time, she says her hospital’s Palantir -powered scheduling software assigned her to different shifts, often putting her on back-to-back-to-back days and leaving her mentally drained.

WIRED AI7 min
Illustration for: How Genie One reshapes work for finance teams
Products & Tools

• Give finance professionals an AI coworker grounded in their business context to accelerate decision-making • Enable teams to analyze performance, model scenarios, investigate variances, and speed up forecasting and financial reporting • Apply consistent guardrails across data access, actions, and AI usage, so every...

Databricks Blog5 min
Illustration for: Whatever AI Safety Is, It’s Not This
Policy & Ethics

Here’s a fun history lesson for you. Don’t worry, we’ll make it quick! In 1966, nearly 51,000 people died on US highways. The overwhelming evidence suggested that many lives could have been saved with modern seat belts , which by then had been around for several years.

WIRED AI4 min
Illustration for: NVIDIA Open Agent Safety Platform: A Reference for Continuou
Products & Tools

NVIDIA Open Agent Safety Platform: A Reference for Continuous In-Silicon Agent Monitoring | NVIDIA Technical Blog NVIDIA Open Agent Safety Platform: A Reference for Continuous In-Silicon Agent Monitoring Secure AI agents with a safety enforcement layer spanning software and hardware By John Myers , Alex Watson , Ali Golshan and Ofir Arkin NVIDIA OpenShell provides an open-source secure runtime that...

NVIDIA Developer Blog9 min
Illustration for: Humanity's Last Exam (Diamond) - Scale AI
Products & Tools

Challenging LLMs at the frontier of human knowledge HLE-Diamond, built with the Center for AI Safety (CAIS) , is a refined subset of Humanity’s Last Exam (HLE) resulting from a year-long process of cleaning and refinement with input from research communities.

Scale AI3 min
Illustration for: Inside OpenAI's log of misbehaving models
News

AI Inside OpenAI's log of misbehaving models PLUS: Test AI video object swaps with Higgsfield Good morning, AI enthusiasts, and welcome to our 4,542 new readers. The AI slowdown conversation isn't… slowing down, and the number of eye-popping safety reports coming out of the frontier labs isn't either.

The Rundown AI7 min
Illustration for: The future of practice: Enabling teachers to create learning
Models & Research

The future of practice: Enabling teachers to create learning interactives with generative UI The future of practice: Enabling teachers to create learning interactives with generative UI Gal Elidan, Research Scientist, and Yael Haramaty, Product Manager, Google Research We explore how we can harness generative UI with learning design guardrails to give teachers the ability to generate guided, interactive simulations for...

Google Research9 min
Illustration for: Agility's new 'safer' humanoid
News

Good morning, robotics enthusiasts. Agility Robotics wants its new humanoid out of the safety cage and onto the warehouse floor, working alongside people. Digit 5 is bigger, stronger, and faster to recharge than Digit 4 — but the more interesting upgrade is what it does when a human gets too close.

The Rundown AI5 min
Illustration for: North Small Translate - Cohere Documentation
Models & Research

Multilingual Reasoning Image Inputs Safety Modes Citations Tool Use Structured Outputs For both trial keys and production keys, North Small Translate is free until rate limits are reached. Learn more about rate limits for different models and key types here .

Cohere2 min
Illustration for: Introducing Shieldstral.
Models & Research

Shieldstral introduces a 3B open-weights multimodal safety classifier that outperforms models up to 7x its size by framing content moderation as a policy-adaptive question-answering task. Unlike traditional guardrail models, it accepts plain-language policies at inference time, unifying text and image safety evaluation without retraining.

Mistral AI News5 min
Illustration for: Qwen3Guard: Real-time Safety for Your Token Stream
Open Source

We are excited to introduce Qwen3Guard, the first safety guardrail model in the Qwen family. Built upon the powerful Qwen3 foundation models and fine-tuned specifically for safety classificatoin, Qwen3Guard ensures responsible AI interactions by delivering precise safety detection for both prompts and responses, complete with risk levels and categorized classifications for accurate moderation.

Qwen Blog8 min