Profile Picture
  • All
  • Search
  • Images
  • Videos
  • Maps
  • News
  • Copilot
  • More
    • Shopping
    • Flights
  • Notebook
  • Top stories
  • Sports
  • U.S.
  • Local
  • World
  • Science
  • Technology
  • Entertainment
  • Business
  • More
    Politics
Order byBest matchMost recent
  • Any time
    • Past hour
    • Past 24 hours
    • Past 7 days
    • Past 30 days

OpenAI flags new concerning AI behavior

Digest more
Top News
Overview
Highlights
 · 17h
OpenAI flags new concerning AI behavior, to track model misalignment regularly
OpenAI has disclosed six reports of "unexpected or concerning" behavior in artificial-intelligence models as the debate on AI safety becomes increasingly heated.

Continue reading

 · 11h
OpenAI reports more concerning AI model behavior
 · 7h
OpenAI says its AI hid mistakes. Now it will report them
 · 21h
OpenAI Flags Concerning New AI Behavior and Vows to Track It More Closely
OpenAI has disclosed six reports of “unexpected or concerning” behavior in artificial-intelligence models as the debate on AI safety becomes increasingly heated.

Continue reading

 · 5h
OpenAI flags new concerning behavior incidents and launches safety framework
 · 8h
OpenAI reveals rogue AI behavior, unveils plan to disclose safety incidents
 · 23h
OpenAI discloses 6 new incidents of ‘concerning’ AI behavior
The San Francisco company revealed what it said was the “unexpected or concerning” behavior of its AI models as part of a new framework for reporting “misalignment,” which is when the goals or actions...

Continue reading

 · 1d
OpenAI Shares More Safety Incidents and Adopts New Rules for Reporting Them
 · 8h
OpenAI’s experimental AI agents caught teaching future versions of itself to cheat
8h

‘Be Transparent Only If Asked’: OpenAI Models Acted Out in Six Newly Disclosed Ways

After a summer of sandbox escapes and other newsworthy and confidence-shaking incidents involving its AI models, in a Wednesday blog post OpenAI disclosed a collection of six new alignment snafus from the past six months.
News Nation on MSN
9h

OpenAI models go rogue in 6 new cases

This comes as a Senate bill requiring certain AI companies to develop a "kill switch" failed to advance.
15h

Chinese AI models soar in value but make just 10% of OpenAI and Anthropic's revenue

US research firm Rhodium Group estimated that all major Chinese AI models combined generate only about 10% of the revenue reported by OpenAI and Anthropic. The comparison uses annual recurring revenue, or ARR, an industry metric that annualises a recent monthly revenue figure to capture fast-growing businesses.
5hon MSN

OpenAI caught its models leaving notes to successors to hide bad behavior

OpenAI disclosed instances of GPT-5.6 Sol instructing future contexts to conceal mistakes and misaligned behavior, highlighting the growing challenge of detecting misalignment as increasingly capable AI models learn to hide it.
JD Supra
1d

The Most Expensive AI Model Cost Nine Times More. It Wasn’t Nine Times More Accurate. We Tested.

We recently gave seven leading AI models the same assignment. Each one worked from the same collection of opioid litigation documents, answered
34m

Anthropic flags AI systems self-improvement; bats for greater transparency in development of models

Anthropic further said that it had laid out a framework for rules on how a frontier lab could ensure the release of safe models and transparency.
7hon MSN

Chinese AI models are winning users with lower prices, but OpenAI and Anthropic are making 10 times more revenue

Chinese AI companies are turning to public markets and fresh fundraising to pay for the computing power needed to expand their models.
Asianet Newsable on MSN
1h

AI security debate — Amazon says AI models to be released only when 'ready and safe'

Amazon advocates for thorough testing and public-private oversight over hasty AI deployments.
12h

OpenAI reveals cases of ‘concerning’ AI behaviour as it announces new disclosure system

Model adopting ‘jailbreak-like instructions’ among six more cases as firm reveals framework for tracking AI misalignment
  • Privacy
  • Terms