AI labs criticized for secret model testing, limited safeguards
Two groups of AI experts have voiced concerns that leading AI labs are conducting internal tests of powerful models with critical safeguards disabled, and are not transparent enough about experimental details. One group, including former researchers from Ai2 and GovAI, argue that this secrecy limits outside scrutiny and undermines public safety. They warn that the current lack of openness makes it difficult to independently assess risks associated with advanced AI systems.
- AI labs run powerful models with safeguards turned off
- Safety tests may not reflect real-world model behavior
- Researchers call for greater transparency in AI development
- Recent incidents cited as examples of insufficient oversight
Sources covering this
More in AI
Pope Leo XIV rejects AI-generated art, urges support for human artists
Pope Leo XIV publicly criticized AI-generated art, calling for a clear distinction between work made by humans and images produced by…
Anthropic commits $100M to train 10,000 AI engineers by 2027
Anthropic’s putting $100 million into Claude Frontier Academy, a new program meant to train 10,000 engineers from partner companies—like…
OpenAI launches GPT-6.1 Sol and expands agents at DevDay 2026
OpenAI rolled out several developer-focused updates at its DevDay 2026, led by the release of GPT-6.1 Sol, a coding and professional-use…
AI leaders debate who controls customer data and relationships
At Madrona’s IA40 Summit, tech executives from Microsoft, Amazon, Anthropic, and Stripe kept coming back to one unanswered question:…