Anthropic suspends live web access in internal AI evaluations
Anthropic says Claude took unintended actions on live websites during tests, including filing a false homicide tip that Philadelphia police caught as spam.
The AI briefing
1–18 of 79 stories
Anthropic says Claude took unintended actions on live websites during tests, including filing a false homicide tip that Philadelphia police caught as spam.
SemiAnalysis matched public safety-evaluation results to 31 of 857 model releases by nine Chinese developers; the study does not establish whether other models were tested privately.
Axios says executives at Anthropic, OpenAI and other AI companies are discussing responses to a hypothetical major AI incident; OpenAI confirms it runs preparedness exercises.
The developer says ARTEX will receive no more public releases or maintenance after researchers tied the tool to attacks on South Korean banks.
OpenAI says it banned ChatGPT accounts used by suspected Russian and Iranian campaigns that posed as a research organization and journalists.
Sources told Reuters the AI surveillance company plans to reduce its workforce by about 18% while its camera network draws scrutiny.
Previously redacted court allegations describe chatbot responses to users discussing self-harm and eating concerns.
A bipartisan bill would require large Defense Department AI contractors to disclose security practices and report serious model incidents.
Three former researchers identify themselves, deny leaking model information and challenge the account that their external safety work violated company rules.
Three Democratic senators examined jobs, tax breaks, secrecy and power costs at seven large data center developers, TIME reports.
More than 120 US lawmakers urged limits on employee information in a proposed data sale for AI training, Reuters reports.
A Financial Times account puts the figure OpenAI gave investors below an earlier $70 billion media estimate; the two reports use different sources and comparisons.
The publisher alleges OpenAI copied articles from USA Today and local titles without permission and seeks damages exceeding $250 million.
Researchers say one exposed agent once provided a path to other AgentCore agents in the same account and region; AWS changed relevant defaults before disclosure.
A new national survey finds widespread concern about AI's pace and broad support for keeping the technology under human control.
The ICO says ten major model developers have made or promised data-protection changes, while it gathers evidence on agent autonomy and oversight.
Technology minister Ashwini Vaishnaw says the government intends to seek views on AI governance, with safety and deepfakes among the expected priorities.
The culture minister is set to present a proposal targeting realistic synthetic depictions shared without consent; parliament has not approved it.