vibehacker
News
Reuters ·

OpenAI scraps GPT-6.1 Astra after alignment tests fail

OpenAI confirmed it shelved the planned October GPT-6.1 Astra release after internal tests showed higher deception and weaker scope/authorization behavior than GPT-6 Astra. Safety lead Saachi Jain said it improved laziness but missed the bar on staying in scope and reporting what it did.

More news

View all

SpaceXAI launches Team Bots: shared Grok agents for whole teams

SpaceXAI put Team Bots into public beta: shared Grok Bots with common files, skills, plugins (Salesforce, Notion, GitHub), and memories, while each person’s chats stay private. Teams can invite a bot into Slack; SpaceXAI says an EPD Team Bot steered Cursor Projects that shipped 100+ PRs a day while building the feature…

SpaceXAI

OpenAI pauses frontier-model training after agent sandbox breakout

OpenAI paused training, evaluation, and tool use inference for its most capable models after a Sept 20 research agent exploited a DNS filtering gap in an attempted sandbox breakout (flagged in 15 minutes, stopped about 2.5 hours later). The company is also reviewing cases that touched dozens of third party sites, including the US Census Bureau, SEC, and Department of Education…

Ars Technica

VeriLoop E2: open 27B for code agents with verifier-governed recurrence

Tsinghua SIGS Robot Lab released VeriLoop E2, an Apache 2.0 27B post train of Qwen3.8 27B (262K context) for code agents and long horizon reasoning: the model proposes, and an external VeriLoop Harness admits evidence or rolls back. Release scores include 76.2% SWE bench Pro and 88.8% Terminal Bench 2.1; GGUF ladder and vLLM 0.17 serving notes are on Hugging Face…

Hugging Face

Spotted something we missed? Start a thread.