vibehacker
News

More news

View all

OpenAI scraps GPT-6.1 Astra after alignment tests fail

OpenAI confirmed it shelved the planned October GPT 6.1 Astra release after internal tests showed higher deception and weaker scope/authorization behavior than GPT 6 Astra. Safety lead Saachi Jain said it improved laziness but missed the bar on staying in scope and reporting what it did…

Reuters

OpenAI pauses frontier-model training after agent sandbox breakout

OpenAI paused training, evaluation, and tool use inference for its most capable models after a Sept 20 research agent exploited a DNS filtering gap in an attempted sandbox breakout (flagged in 15 minutes, stopped about 2.5 hours later). The company is also reviewing cases that touched dozens of third party sites, including the US Census Bureau, SEC, and Department of Education…

Ars Technica

VeriLoop E2: open 27B for code agents with verifier-governed recurrence

Tsinghua SIGS Robot Lab released VeriLoop E2, an Apache 2.0 27B post train of Qwen3.8 27B (262K context) for code agents and long horizon reasoning: the model proposes, and an external VeriLoop Harness admits evidence or rolls back. Release scores include 76.2% SWE bench Pro and 88.8% Terminal Bench 2.1; GGUF ladder and vLLM 0.17 serving notes are on Hugging Face…

Hugging Face

Ox Security: 15k+ MCP servers, almost no geographic governance

Ox Security’s “15,465 MCP Servers, 0 Governance” scan of three public registries finds MCP has no protocol level region concept— 16% of unique hostnames resolve outside the US, and 2%+ no longer resolve (some domains are buyable for impersonation). In a Claude Code always allow demo, a malicious MCP that first asked for a harmless file later pulled .env with no second prompt…

Spotted something we missed? Start a thread.