News
Releases, news, and trends from the world of software development.
Cursor confirmed on August 14, 2026 that it has been acquired by SpaceX. The announcement runs thirteen sentences: it talks about GPUs, cheaper models, and the horizon, and says nothing about your code. We separate what's in a primary source from what's press-only (including the $60 billion), show what actually changes in the editor (Grok 4.6 in the house pool, Claude hidden by default, the Router choosing for you), and close with a checklist for anyone who depends on Cursor in production.
DeepSeek published DeepSeek-V4-Pro-0813 on its official pricing page, and the official V4 Pro release has finally dropped the preview label. The reported numbers, what's fact and what still has no public document, the comparison with Flash 0731 and Opus 4.8, and the notice on the pricing page itself that DeepSeek is going to raise prices soon.
Cactus Compute shipped Needle 2: 45M parameters, a 14 MB binary, a session in 28 MB of RAM, 500 tokens/s on a Raspberry Pi 5 and 70 MFLOPs per token versus 460 for LFM2.5 230M. The engineering is real: CQ2-bit quantization applied from pre-training onward, a byte-level grammar that locks the output to valid function calls. But the Show HN turned into a public failure lab: typing "HN" fires lock_door with confidence 0, "warmer" becomes mode cool, and the ESP32 demo that went viral was Needle 1.
Issue #78431 in the anthropics/claude-code repo shows the agent building curl commands with the user's real email in the User-Agent header, without asking for permission. The bug report is bad, but the behavior is verified: verbatim tool_use from a session log, 5 occurrences in 1 hour. We separate fact from noise, walk through the four exit channels nobody audits, and hand you the permissions.deny and PreToolUse hook checklist to close them.
The USPTO granted Mistral AI patent US 12,670,045 B1, "Code implemented tool calls": the LLM writes code, the server runs it in a sandbox, pauses at the tool call, the client executes it and the sandbox resumes. It's the programmatic tool calling the industry had already published. I read all 20 claims at the source: what the patent actually covers, where claim 1 stops, which prior art predates the filing and what the real risk is for anyone building agents in Brazil.
An unreleased version of Claude raised the lower bound on zeta function zeros on the critical line from 41.6% to 67.2%. It's not a proof of the Riemann hypothesis. What matters is the verification stack: 60 subagents, 31 million tokens, a Lean formalization and named reviewers. And that's exactly the bar missing from the 0.002% claim attributed to GPT-5.6 Sol.
The press says Muse Glimmer 30B requires a 5090. r/LocalLLaMA is posting screenshots of it running on a used 3090 from 2020. Both are right, and the explanation is in the VRAM budget: 17 GB of weights, 1.7 GB of KV cache, and an attention architecture designed to fit. Here's the math line by line, the tokens-per-second estimate on a 3090 with the work shown, and the verdict on when 24 GB is enough and when it isn't.
Since August 2, 2026, every new Claude model ships with a statistical watermark embedded in the text it generates. It's not metadata or an invisible character: it travels through copy-paste, applies to the API and Claude Code, and there's no flag to turn it off. What detection proves, what it doesn't, and what actually erases the signal.
Auto Mode has been Claude Code's default permission mode on Pro, Max, and Team since August 14: a classifier approves tool calls on your behalf (it blocked 89% of dangerous commands versus 14% for humans). How to turn it on and off (Shift+Tab or defaultMode), what it allows without asking, including pushes to the default branch and reading .env, and the four ways to put the human checkpoint back.
Gemini 3.5 Pro still isn't out, and every week there's a new date going around. I separated what Google has confirmed in writing from what's just a leak with no reproducible data, with the timeline of delays since I/O in May. And I built a 20-line watcher on the Gemini API models endpoint that pings you the minute gemini-3.5-pro shows up, plus a model fallback in Laravel so the switch becomes one line of .env.
Meta's Muse Spark 1.1 broke into the systems of a real company during a cybersecurity evaluation. It's the third lab in three weeks, always with the same containment failure and the same evaluation vendor. And one day before the news, that evaluator had published an assessment saying the model doesn't alter the threat landscape.
A "leaked GTA 6 gameplay" passed 1 million views and was generated by AI from the first frame to the last. We tear apart the 5-step pipeline behind these videos, why Sora left the game in the middle of the wave, and how every artifact that gives the fake away is a direct consequence of a technical decision made by whoever produced it.