:

INDEPENDENT AI TESTING LAGS BEHIND MODEL ADVANCES

AI DESK1 MIN READ
MON, SEP 14, 2026

■ AI-SUMMARIZED FROM 1 SOURCE ▸ TIMELINE

Investment in AI safety testing has not kept pace with rapid improvements in model capabilities, according to Vals AI co-founder Rayan Krishnan. His firm independently evaluates models from OpenAI, Anthropic, and others.

Krishnan's independent testing lab has identified early signs of recursive self-improvement in current AI systems, though existing models remain significantly behind top human researchers in capability. The gap between model advancement and safety evaluation presents a growing challenge for the AI industry. As companies release increasingly powerful systems, the infrastructure to thoroughly test and understand their behavior has lagged behind. Vals AI conducts third-party evaluations of AI models to identify potential risks and limitations before deployment. This independent oversight addresses concerns that companies evaluating their own systems may miss critical safety issues. The company's findings suggest that dedicated investment in AI testing infrastructure is needed to keep pace with development cycles. Current evaluation methods may not be sufficient to catch emerging behaviors in more advanced systems. Krishnan's comments highlight a core tension in AI development: rapid capability improvements outpacing the ability to safely evaluate and constrain those capabilities.

■ SOURCES

Bloomberg Tech

■ SUMMARY WRITTEN BY AI FROM THE LINKS ABOVE

■ MORE FROM THE AI DESK

Prominent voices in artificial intelligence are pushing for slower development to allow safety measures to catch up with increasingly powerful models. The debate is gaining traction in Washington and on Wall Street, where AI-linked stocks face pressure amid investor concerns about the implications for spending growth.

JUST NOWAI Desk

Google's Gemini AI can now assist users in organizing cluttered Google Drive files and folders. The feature suggests new organizational structures for files, photos, and folders.

JUST NOWAI Desk

Cohere CEO Aidan Gomez has pushed back against proposals for major AI labs to coordinate on safety standards, arguing the approach could entrench dominant players like OpenAI and Anthropic. Gomez called for broader stakeholder involvement in shaping AI regulations instead.

1H AGOAI Desk

A new study shows that the rise of artificial intelligence has coincided with declining initial earnings and employment rates for college graduates in AI-exposed fields. The trend raises concerns about long-term career trajectories.

1H AGOAI Desk

■ SUBSCRIBE TO THE DAILY BRIEF

ONE EMAIL, 5 STORIES, 06:00 UTC. UNSUBSCRIBE ANYTIME.