Yet another concerning result on Astra's no-CoT capabilities

·LessWrong··

This is a research update for an on-going replication of no-CoT[1] evals done as part of the Second Look Fellowship. In following posts, we will run more comprehensive replications of previous work and release open source tooling for no-CoT eval elicitation. Code can be found here.tl;drWe replicate experiments from Greenblatt 2025 and Greenblatt 2026 on GPT-6-Astra, on the same items and protocol as our previous update on Fable 5, Opus 5, Opus 4.5, and GPT-5.6-Sol, plus Gemini 3.1 Pro, Kimi k3, ...

Read full article →

Related Articles

Why are AI agents lying, cheating and coordinating?
jonifico · Hacker News · 1d ago
JetKVM Mini
taubek · Hacker News · 19h ago
I'm being cyberattacked by Tesla, Inc
robinpie · Hacker News · 9h ago
google.com/goto: Google's anti-scraping update
1e1a · Hacker News · 2d ago
Revolut confirms customer data breach through fake government requests
tdrz · Hacker News · 17h ago