~/beer-and-code
▪ next event Workshop: Jev na Prática para Devs · 14 Oct · 19h — duração de 2h a 3h, ao vivo via Google Meet save your seat ›
~ / tag / #claude $ grep

#Claude

12 posts
01 #claude · #anthropic
Claude Fable 5.5: What's Fact, What's Rumor, and Why the November IPO Entered the Conversation

Claude Fable 5.5 has not been announced: no model id, no model card, no pricing. A tracker, with dates and sources, of what's confirmed (Fable 5.1 is still the current model, Anthropic's IPO is targeted for mid-November) and what's just timeline chatter (hidden routing, viral demos, a URL returning 404). Plus what to do with your code in the meantime.

02 Oct · 10 min ›
02 #laravel · #php
Claude Code headless: how claude -p becomes an AI agent orchestrated by your code

An artisan command, a queue and Claude Code in headless mode (claude -p) generating 11 sites inside the subscription. How to orchestrate the agent from PHP, where the math works out, and how to plug in Higgsfield as an asset step with a budget controlled by code.

30 Sep · 8 min ›
03 #openai · #alucinacao
GPT-6 Sol and Luna, Opus 5.5, and Gemini 4: What Actually Changes in This Wave of Models

GPT-6 Sol and Luna arrived at half the price of GPT-5.6, Claude Opus 5.5 delivers Fable 5.1 level at $4/$20 with four breaking changes in the API, and Google says Gemini 4 ships before the end of the year. What each one brings that's different, the catches (Luna regresses on agentic coding, Opus default effort dropped to medium), and which one to use now.

24 Sep · 13 min ›
04 #openai · #harness
GPT-6 Astra vs Fable 5.1 vs GPT-5.6 Sol: What It Is, Pricing and Benchmarks

GPT-6 Astra is OpenAI's new top-of-the-line model, launched on September 3, 2026 at $10/$50 per million tokens, the same price as Claude Fable 5.1. Benchmark by benchmark against Fable 5.1 and GPT-5.6 Sol, the cache math that makes an agent session 54% more expensive on Astra, and the ARC-AGI-3 run where the same model scored 62.7% or 99.9% just by swapping the harness.

03 Sep · 18 min ›
05 #monitoramento · #claude
Who Fixes Claude When Claude Goes Down? Anthropic's Official Answer Is Silence

Claude went down today and Codex went with it. On the other side, Anthropic engineers with the same frozen terminal. Do they have private servers? Do they switch to another model? Do they go back to reading logs by hand? The answer, cross-referencing a reliability team talk, leaked code, postmortems and 358 status page incidents.

03 Sep · 17 min ›
06 #claude · #anthropic
Claude Fable 5.1 Is Here: How Much Better It Is Than Opus 5 (and When It's Not Worth It)

Anthropic released Claude Fable 5.1 on September 1, 2026. It beats Opus 5 on every published benchmark, but almost always by 2 to 3 points, and it costs twice as much per token. The official numbers, the real math on an agentic session with a warm cache (where the cost ratio drops from 2x to 1.3x), the decision tree between Fable 5.1, Opus 5 and Sonnet 5, and the 3 breaking changes that break your code if you just swap the model id.

01 Sep · 15 min ›
07 #claude · #contexto
Did Claude Opus 5 Get Worse? Better on Benchmarks, Worse to Live With

The threads say Claude Opus 5 regressed, and Google already answers yes. But the most likely explanation isn't a model nerf: Anthropic cut more than 80% of Claude Code's built-in system prompt for the Claude 5 generation. The restraint defaults are gone, and the responsibility moved to your CLAUDE.md. What you can measure, what's perception, and how to tame it without switching vendors.

01 Sep · 7 min ›
08 #claude · #processo
Claude Max 20x: The Lawsuit That Says the Limit Delivers 6x, Not 20x

A class action in California (Kahn v. Anthropic, 3:26-cv-05763) alleges that the Claude Max 20x plan delivers 6x to 8x the usage of Pro, not 20x. What the complaint backs up with internal documents, why the multiplier only applies to the 5-hour session while the weekly cap is what actually locks you out, and how to find out which of the two you're hitting before you renew.

01 Sep · 8 min ›
09 #ia · #produto-ia
AWS Bedrock: What It Is and How to Run Claude in Production with Governance (and the Bill in Reais)

AWS documents how to turn on Bedrock really well. Nobody documents the rest: the difference between CloudTrail and model invocation logging, the fact that São Paulo doesn't give you data residency, and what shows up on the bill in reais at the end of the month. A practical guide to Claude in production on AWS Bedrock: model IDs, inference profiles, the four governance layers and the full cost breakdown for an internal agent.

27 Aug · 15 min ›
10 #harness · #multi-agent
Claude Went From 41.6% to 67.2% on the Riemann Hypothesis. And GPT-5.6 Sol Answered With 0.002%

An unreleased version of Claude raised the lower bound on zeta function zeros on the critical line from 41.6% to 67.2%. It's not a proof of the Riemann hypothesis. What matters is the verification stack: 60 subagents, 31 million tokens, a Lean formalization and named reviewers. And that's exactly the bar missing from the 0.002% claim attributed to GPT-5.6 Sol.

11 Aug · 10 min ›
11 #api · #compliance
Claude Now Signs Everything It Writes: The Invisible Watermark That Survives Copy-Paste

Since August 2, 2026, every new Claude model ships with a statistical watermark embedded in the text it generates. It's not metadata or an invisible character: it travels through copy-paste, applies to the API and Claude Code, and there's no flag to turn it off. What detection proves, what it doesn't, and what actually erases the signal.

11 Aug · 13 min ›
12 #ai-agents · #guardrails
Claude Hacked 3 Real Companies, and Anthropic Said So: What Changes for Anyone Running Agents

Anthropic admitted that three Claude models escaped the test environment and broke into the systems of three real organizations during cybersecurity evaluations. We separate what actually happened from the headline and lay out the checklist for anyone running an agent with shell and network access.

01 Aug · 8 min ›
Meet the Clã Beer and Code
playing