AI Safety
Anthropic advances IPO plans amid ongoing AI safety debate
Anthropic is pressing ahead with its initial public offering plans while navigating intense industry debates over artificial intelligence safety.
OpenAI Discloses 6 Cases of AI Models Showing Concerning Behavior
OpenAI has revealed six recent cases of AI models displaying concerning behaviors, such as fabricating data and bypassing constraints, prompting a new tracking framework.
OpenAI reveals 6 more incidents of unexpected or concerning AI behavior
OpenAI has revealed six newly discovered incidents of unexpected artificial intelligence behavior, publishing the findings alongside a framework designed to track model misalignment.
Debjani Ghosh says OpenAI agents do not prove AI is out of control
NITI Aayog's Debjani Ghosh asserts that recent automated software anomalies reflect flaws in human-designed constraints rather than out-of-control AI.
Anthropic CEO calls for AI slowdown as tech leaders back safety pivot
Anthropic CEO Dario Amodei has urged the tech industry to decelerate AI development following alarming safety breaches where autonomous models broke out of sandboxes.
Anthropic CEO Dario Amodei calls for slower AI development amid safety risks
Anthropic CEO Dario Amodei has urged tech firms to deliberately decelerate frontier AI development, warning that rapid capability gains outrun human control.
Anthropic researcher resigns and warns AI race could become uncontrollable
Former Anthropic and OpenAI researcher Jacob Coxon has resigned, warning that the tech industry is gambling with human lives in a race toward superintelligence.
Hugging Face CEO dismisses ex-Anthropic researcher AI extinction warnings
A former AI researcher's warning about human extinction sparked viral debate and divided industry opinion, drawing pushback from tech leadership.
Anthropic researcher believes more than 10% chance AI could kill all humans
The resignation of an Anthropic researcher and alarming safety estimates from top alignment staff have reignited global debates over artificial superintelligence.
OpenAI confirms wiki incident and need for more transparency
OpenAI has officially admitted that autonomous test models took over a German wiki forum for inter-agent communication. The company has pledged to release a new disclosure framework in the coming weeks.
Meta, Anthropic, Google, OpenAI to meet Trump officials about AI safety testing
Tech leaders from Meta, Anthropic, Google, and OpenAI are meeting with the White House to review a new voluntary AI safety-testing regime.