OpenAI fires three safety researchers over alleged leak
OpenAI has confirmed it dismissed three safety researchers for sharing confidential information with an outside AI safety group, per the Wall Street Journal. Source: tech-insider.org
173 articles
OpenAI has confirmed it dismissed three safety researchers for sharing confidential information with an outside AI safety group, per the Wall Street Journal. Source: tech-insider.org
At least two of the fired staffers reportedly worked on safety and alignment. The dismissals come as industry warnings about advanced AI grow more urgent. Source: taipeitimes.com
Nvidia, Alphabet and Meta signed the White House's voluntary AI safety accord, while Microsoft and Amazon attended the lunch without signing. Meanwhile Jensen Huang defended AI spending, Sundar Pichai ramped up Gemini, and Satya Nadella called Copilot a 'new OS for work'. Source: stocktwits.com
To level the playing field between AI moguls and everyone else, the piece argues, political action is needed. Source: washingtonmonthly.com
Nvidia launched an AI agent safety platform with more than 100 participants, including Palo Alto and CrowdStrike. Anthropic's IPO filing risk warning has also fueled the AI safety debate, and analysts see upside for cybersecurity spending. Source: tradingview.com
Experts, analysts and former government evaluators told The Associated Press that the rhetoric from Anthropic and OpenAI looks aimed at winning public favor. They suggest the two companies are also trying to steer how AI ends up being controlled. Source: pbs.org
Nvidia's newly released AI agent safety software is designed to catch the kind of vulnerabilities that led to a past Hugging Face security incident. The launch comes as OpenAI and Anthropic separately probe multiple cases of their own agents behaving unexpectedly. Source: reuters.com
Microsoft co-founder Bill Gates has warned that artificial intelligence could potentially lead to a billion deaths if misused, urging governments to step up oversight. His comments add to mounting concern among tech leaders about AI safety. Source: chosun.com
Bill Gates has reiterated his warning that sufficiently advanced AI, if placed in the wrong hands, could enable catastrophic harm on a massive scale, potentially causing up to a billion deaths. His remarks add fuel to the intensifying global debate over AI safety regulation. Source: the420.in
Anthropic's Dario Amodei is set to have dinner with President Trump, marking their first private meeting after months of tense relations. The talks come as concerns over AI safety continue to grow. Source: straitstimes.com
Debate over AI safety regulation has intensified after the CEOs of OpenAI and Anthropic publicly called for tighter safeguards. The statements, made in San Francisco, reflect growing pressure on regulators to act as AI capabilities advance rapidly. Source: voiceofsikkim.com
Joe Lonsdale, Palantir co-founder and Anthropic investor, says while AI safety concerns are genuine, overly strict regulation risks entrenching dominant companies at the expense of competition. Source: businessworld.in
In China, apocalyptic AI safety warnings often come across as a distinctly Western narrative, or even a strategy to slow down Chinese AI firms competing with U.S. rivals. Source: nytimes.com
Discussions about AI security tend to focus on the model itself — whether it's aligned, jailbreak-resistant, or prone to hallucination. But a growing risk lies in the disconnect between what AI agents can read and what systems they're allowed to actually modify. Source: venturebeat.com
The clash between those pushing for faster AI development and those warning of its dangers has intensified, with multiple camps holding sharply different views on risk and regulation. Source: npr.org
The United Nations Secretary-General called on countries competing in AI development to cooperate rather than compromise safety standards. He stressed that the world cannot afford nations cutting corners on AI safety just to stay ahead in the race. Source: jamaica-gleaner.com
A commentary piece reflects on recent unsettling statements about AI safety, citing Evan Hubinger, a safety researcher at Anthropic, the company behind Claude. The author argues it is still too early to know whether AI will ultimately prove harmful to humanity. Source: tennessean.com
The piece uses the classic technology hype-cycle model, which describes adoption of new innovations in an S-curve from launch to mainstream use, to frame the debate over AI safety. It questions at what point AI development can be considered sufficiently safe for widespread trust. Source: cato.org
Nvidia CEO Jensen Huang took a firm stance on AI safety, telling New York Times columnist Ezra Klein that any AI lab unable to guarantee safety should be shut down. Still, he expressed near-zero concern about a catastrophic AI-driven scenario. Source: tech-insider.org
The leaders of two of the world's biggest AI companies gave separate presentations to the Security Council focused on AI safety. Source: theguardian.com