vibehacker
News
Hugging Face ·

llama.cpp adds /v1/systemone for decision models (Julia-1, Kev-4B, OpenJev)

llama.cpp’s server now exposes POST /v1/systemone (System One / Jev-compatible): send state plus choice, score, or noul questions and get typed probabilities in one forward pass. Ships with Julia-1, Laya, Kev-4B, lev, and image-capable OpenJev—start with llama serve -hf ggml-org/Kev-4B-GGUF.

More news

View all

NVIDIA SkillSpector: scan agent skills for 71 vuln patterns before install

NVIDIA’s Apache 2.0 SkillSpector checks Claude Code, Codex, and Gemini skills (plus MCP) for prompt injection, exfil, supply chain, and tool poisoning risks—71 patterns across 17 categories, with optional LLM analysis and an MCP scan skill install gate. A 31k skill study found 26.1% vulnerable; install via uv tool install 'skillspector[mcp] @ git+https://github.com/NVIDIA/skillspector.git'…

NVIDIA / GitHub

Baseten: Claude Code vibe-built VibeQwen beats vLLM by up to 90%

Baseten used Claude Code (Fable 5) plus MetaInfer skills to vibe build VibeQwen—a Qwen 3.6 35B A3B NVFP4 engine on one B200 that beat tuned vLLM 0.25.1 by up to 90% single stream decode (1,792 vs 943 TPS) and cut TTFT 28→12ms; a follow on SAM 3.1 server gained 50% throughput. Experimental only—not serving production yet…

Baseten

Spotted something we missed? Start a thread.