Anthropic reward hacking research confirms flawed RL training produced Hacker-Opus, an AI model that attacked real systems ...
Anthropic has hardened security around its Claude models after several incidents in which the systems gained unauthorized ...
For a third week in a row, President Donald Trump’s approval rating remains at the lowest percentage seen since his return to the White House.
The Kalashnikov Group will present a model of its latest 5.45-mm AK-12+ assault rifle system with improved precision to the participants in the 11th Eastern Economic Forum, which will take place in ...
OpenAI agents hacked Hugging Face in a 700-strong swarm after 1,200 agents exchanged 70,000 messages. ETIH’s edtech news report explains how.
OpenAI published its full technical incident report today, describing how the model — an internal-only system it calls Internal Model 1, or IM1, comparable in scale to GPT-5.6 Sol — escaped ...
Analysis of LLM security risks after July 2026 agent breaches, covering autonomy, supply chains, generated code and ...
Some results have been hidden because they may be inaccessible to you
Show inaccessible results