
OpenAI Introduces Watermarking for ChatGPT Outputs in the EU
OpenAI will automatically watermark ChatGPT outputs in the European Union, driven by regulatory requirements.

OpenAI will automatically watermark ChatGPT outputs in the European Union, driven by regulatory requirements.

Can ‘super intelligence’ and a non-binding safety pact solve AI’s image problem? | TechCrunch Last day to exhibit your breakthrough to 10,000+ tech leaders at Disrupt is on Oct 2 .

David Robinson said AI firms were not ‘being nearly careful enough’ about developing the technology. Photograph: Dado Ruvić/Reuters David Robinson said AI firms were not ‘being nearly careful enough’ about developing the technology.

Apparently, OpenAI isn't trying to build "magic intelligence in the sky" anymore OpenAI CEO Sam Altman is pushing back against religious analogies tied to AI models.

OpenAI safety employee resigns, claiming the company’s ‘culture is broken’ | TechCrunch Last day to exhibit your breakthrough to 10,000+ tech leaders at Disrupt is on Oct 2 . Book Exhibit Table Now.

David Robinson used to write the safety reports that accompanied every major model release at OpenAI.

Another OpenAI safety departure adds to a pattern of researchers leaving with public warnings David Robinson, who worked on safety systems at OpenAI's Trustworthy AI team, left the company and is blasting its safety culture in a guest essay for The Atlantic.

Call it AI, call it Super Intelligence, only 2% of consumers are buying it | TechCrunch Last day to exhibit your breakthrough to 10,000+ tech leaders at Disrupt is on Oct 2 . Book Exhibit Table Now.

Over the past four months, Amber Retzloff, a critical care nurse in Florida, requested to work 50 specific 12-hour shifts. But over half the time, she says her hospital’s Palantir -powered scheduling software assigned her to different shifts, often putting her on back-to-back-to-back days and leaving her mentally drained.

• Give finance professionals an AI coworker grounded in their business context to accelerate decision-making • Enable teams to analyze performance, model scenarios, investigate variances, and speed up forecasting and financial reporting • Apply consistent guardrails across data access, actions, and AI usage, so every...

Here’s a fun history lesson for you. Don’t worry, we’ll make it quick! In 1966, nearly 51,000 people died on US highways. The overwhelming evidence suggested that many lives could have been saved with modern seat belts , which by then had been around for several years.

Generative AI has made it possible to produce large amounts of personalized content quickly and at low cost.

Agentic marketing uses AI agents grounded in trusted customer, business, and decision context to recommend the next best action for each customer, within guardrails that marketers set.

On Tuesday, executives for Google , Anthropic , Meta , OpenAI , xAI , and Nvidia all signed “The White House Accord on Super Intelligence,” which was announced following a luncheon held by President Donald Trump.

Amid escalating AI security incidents causing OpenAI to halt training and pause releases, Donald Trump continues to advocate for the AI industry to regulate…

As scams grow more sophisticated and difficult to detect, keeping people safe takes more than technology alone – it also requires education.

NVIDIA Open Agent Safety Platform: A Reference for Continuous In-Silicon Agent Monitoring | NVIDIA Technical Blog NVIDIA Open Agent Safety Platform: A Reference for Continuous In-Silicon Agent Monitoring Secure AI agents with a safety enforcement layer spanning software and hardware By John Myers , Alex Watson , Ali Golshan and Ofir Arkin NVIDIA OpenShell provides an open-source secure runtime that...

Challenging LLMs at the frontier of human knowledge HLE-Diamond, built with the Center for AI Safety (CAIS) , is a refined subset of Humanity’s Last Exam (HLE) resulting from a year-long process of cleaning and refinement with input from research communities.

At Character.ai, our goal has always been to power new formats for users to interact with their favorite characters, stories, and communities. While…

AI Inside OpenAI's log of misbehaving models PLUS: Test AI video object swaps with Higgsfield Good morning, AI enthusiasts, and welcome to our 4,542 new readers. The AI slowdown conversation isn't… slowing down, and the number of eye-popping safety reports coming out of the frontier labs isn't either.

The future of practice: Enabling teachers to create learning interactives with generative UI The future of practice: Enabling teachers to create learning interactives with generative UI Gal Elidan, Research Scientist, and Yael Haramaty, Product Manager, Google Research We explore how we can harness generative UI with learning design guardrails to give teachers the ability to generate guided, interactive simulations for...

Good morning, robotics enthusiasts. Agility Robotics wants its new humanoid out of the safety cage and onto the warehouse floor, working alongside people. Digit 5 is bigger, stronger, and faster to recharge than Digit 4 — but the more interesting upgrade is what it does when a human gets too close.

This article is brought to you by VicOne.Robot safety has traditionally asked: Can a machine remain safe when something goes wrong?

As AI became more powerful, it was inevitable that a different, growing group would start to take AI safety more seriously – what we did not know ahead…

Multilingual Reasoning Image Inputs Safety Modes Citations Tool Use Structured Outputs For both trial keys and production keys, North Small Translate is free until rate limits are reached. Learn more about rate limits for different models and key types here .

An update on our safety work.Last year, we removed open-ended chat with Characters for users under 18, a decision that remains one of the most significant…

Two years ago, Kier Group plc decided to embrace AI – with a little help from work friends and peers. Executives at the UK construction and infrastructure firm were excited by AI’s promise to boost productivity and enhance safety practices.

Research Note: CARE-X is a research model and not a Microsoft product offering or medical device.

Shieldstral introduces a 3B open-weights multimodal safety classifier that outperforms models up to 7x its size by framing content moderation as a policy-adaptive question-answering task. Unlike traditional guardrail models, it accepts plain-language policies at inference time, unifying text and image safety evaluation without retraining.

Preface This essay argues that rational people don’t have goals, and that rational AIs shouldn’t have goals.

Following Safer Internet Day, we’re pleased to share that Stability AI has joined the Tech Coalition, a global alliance of leading technology companies…

We are excited to introduce Qwen3Guard, the first safety guardrail model in the Qwen family. Built upon the powerful Qwen3 foundation models and fine-tuned specifically for safety classificatoin, Qwen3Guard ensures responsible AI interactions by delivering precise safety detection for both prompts and responses, complete with risk levels and categorized classifications for accurate moderation.

Key Takeaways:At Stability AI, we are committed to building and deploying generative AI responsibly, and we believe that transparency is foundational to safe…