vibehacker
News
Startup Fortune ·

MiniMax quietly ships M3.1-Flash-Preview inside MiniMax Code

MiniMax turned on M3.1-Flash-Preview inside MiniMax Code on Sept 27 with no model card, public API listing, or pricing—just a 1M-token context and five reasoning-effort tiers (low through max). Days earlier, community fingerprinting linked OpenRouter’s free stealth/space-bunny-alpha preview to the same MiniMax family; MiniMax has not confirmed the match.

More news

View all

Prismor: open-source runtime guard for Claude Code, Codex, and Cursor

Prismor (formerly Immunity Agent) is a self hosted Apache 2.0 control plane that sits between Claude Code, Codex, Cursor, or Windsurf and every tool call, allowing, warning, or blocking by policy before execution. It wraps npm/pip with supply chain scoring, cloaks secrets in payloads, and ships an MCP Gateway that policy checks calls and injection scans responses; pip install prismor plus prismor setup wires named postures from lightweight dev safe to airgapped…

byteiota

Whiteboard: MIT canvas where agents diagram their changes for review

YC W26’s /dev/fast open sourced Whiteboard, a local desktop canvas where Claude Code, Codex, and similar agents draw sequence and ER diagrams of their work next to an AST aware Rust diff viewer and Decision Log. Diagram nodes jump into a bundled Code OSS editor; installs cover macOS, Linux, and Windows at install.dev.fast ( 1.9k GitHub stars)…

GitHub

llama.cpp prompt-lookup drafting made up to 42× faster

Hayder Tirmazi’s Sept 26 write up speeds llama.cpp’s n gram prompt lookup drafting up to 42× and cuts peak memory 2.6× via reference reads, dense outer maps, sorted vector followers, and a Lemire constmap for the static cache—acceptance rate unchanged. Gains concentrate on repetitive drafting (code edits, structured output), not general chat…

Hayder Tirmazi

BetterWebSearch MCP: keyless search with ~86% smaller research payloads

BetterWebSearch is an MIT licensed local MCP server that defaults to keyless DuckDuckGo search, escalates extraction from HTTP to structured data to Playwright only when needed, and returns cited research passages. On the author’s 12 question bench, web research cut returned text 86% versus open then extract while retaining 33/34 factual markers…

DEV Community

Fireworks Ember-1 matches Kimi K3 quality with ~40% fewer tokens

Fireworks Research’s Ember 1, trained on Kimi K3, cuts reasoning tokens 35–50% while holding coding and agent quality on SWE bench, Terminal Bench, and live customer A/B tests. It’s a research preview serverless option beside K3, aimed at multi turn agent workloads where long traces dominate cost…

Fireworks

Spotted something we missed? Start a thread.