vibehacker
News
GitHub Changelog ·

Copilot CLI's /model picker now finds your local Ollama models and switches to them mid-session

Starting in Copilot CLI 1.0.94-0 (announced Oct 7), /model lists tool-calling, streaming-capable models from a running Ollama instance next to the cloud ones, and you can add one for the current session without restarting. Picking a local model doesn't turn on offline mode or disable telemetry, so set COPILOT_OFFLINE=true if you want that.

More news

View all

Codync turns Claude Code, Codex, and 40 more coding agents into bots you message from your phone

Shown on HN Oct 7, the MIT licensed project runs one Rust binary, codync host , on your Mac, Windows, or Linux machine that drives any agent in the ACP registry with the plans you're already signed into, and you chat with each bot (its own folder, memory, and approval rules) from an iPhone app, desktop app, or terminal. Remote access goes through a free end to end encrypted Cloudflare relay, and bots can ask each other for help, like a Claude bot requesting a Codex review…

Show HN / GitHub

Liquid AI open-sources d1-3B, a decision model that answers in 8 ms on an RTX 4090 with zero output tokens

Released Oct 7 on Hugging Face alongside the experimental d1 omni 600M (text plus image or audio), d1 3B returns a typed answer in one forward pass instead of generating text, scoring 48.57 on the Decision Index v0.2.1 public split, which Liquid says matches Decider 35B A3B at 12x smaller. It has day one llama.cpp support and runs a question in 16 to 50 ms on Jetson boards, so it's a cheap local option for gating, routing, or moderation steps in an agent loop…

Liquid AI Blog

Google ships an official agent skill and gcloud commands so coding agents read live Google docs instead of scraping

Announced Oct 7, npx skills add google/skills skill retrieving developer knowledge teaches Claude Code, Cursor, Copilot, or Antigravity to pull fresh Markdown docs for Google Cloud, Firebase, Android, and more through the Developer Knowledge MCP server, falling back to curl on the REST API. There's also a new gcloud developer knowledge surface (try piping an error trace into answer query ) plus client libraries in seven languages…

Claude Haiku 5.5 lands at $0.10/$0.50 per 1M tokens, 90% under Haiku 4.5 for prompts below 100K

Released Oct 7 for subagents and high volume work, it scores 72.4% on OSWorld 2.1 (offline subset) and 39.2% on Terminal Bench 4.0 vs 16.4% for GPT 6 Luna, and it's the first Haiku with adjustable effort (medium by default). Prompts over 100K tokens cost $0.50/$2.50, and since it uses slightly more tokens per text Anthropic puts the average saving near 75%; it's already live in GitHub Copilot and Amazon Bedrock…

Anthropic

Vercel puts stealth reasoning model Glyph Cluster free on AI Gateway for coding agents

Free during its stealth period for Pro and Enterprise teams with purchased AI Gateway credits, stealth/glyph cluster is a text only reasoning model for coding and long context work that you can pick in Claude Code, Codex, or Cursor after running vercel ai gateway setup . It has function calling but no structured outputs or image input, and there's no zero data retention, so your prompts and responses may be used for training…

Vercel Changelog

Spotted something we missed? Start a thread.