~/beer-and-code
▪ next event Workshop: Jev na Prática para Devs · 14 Oct · 19h — duração de 2h a 3h, ao vivo via Google Meet save your seat ›
~ / tag / #ia $ grep

#Ia

31 posts
25 #ia · #llm
Qwen 3.8 Max (2.4T): Benchmark vs Kimi K3, Pricing and Open Weights

Alibaba dropped Qwen 3.8 Max, a 2.4-trillion-parameter MoE it claims is the second-best model in the world, behind only Fable 5. With no public benchmark at launch, the only independent test scored it 80/100 against Kimi K3's 83. Here: what it is, what it costs (Token Plan from $6 to $68), how to access it via API, and where each model in the Qwen 3.8 family fits.

20 Jul · 9 min ›
26 #ia · #llm
Kimi K3 Pricing: $3/$15 per Million Tokens, Is It Free? Benchmarks vs Claude

Kimi K3 costs $3 input and $15 output per million tokens on the API (the chat at kimi.com is free), 1.7x cheaper than Claude Opus 4.8, not the "5x" your timeline is claiming. We did the honest pricing math, separated the verifiable benchmarks from the hype, and show where it beats Claude (and where it doesn't).

17 Jul · 10 min ›
27 #openai · #ia
GPT-5.6 Sol vs Terra vs Luna: Pricing by Tier and Which to Use

What's the difference between GPT-5.6 Sol, Terra and Luna? Same generation, three tiers: Sol is the agentic top end ($5/$30 per 1M tokens), Terra is the middle ground ($2/$12) and Luna is the cheap, fast one ($0.20/$1.20). With the pricing in effect since the July 30 cut, here's when each one is worth it and when Sol is just wasted budget.

26 Jun · 11 min ›
28 #openai · #ia
Claude Code vs Codex: Codex Wins the Terminal by 13 Points, Claude Wins the Hard Repo by 10

Claude Code or Codex? The answer comes with numbers: Codex opens a 13-point lead on Terminal-Bench (82.7% vs 69.4%) and Claude Code opens a 10-point lead on SWE-bench Pro (69.2% vs 58.6%), which is the benchmark for real multi-file problems. On SWE-bench Verified they tie. Here is the verdict by scenario, the real cost per dev, and the criterion that matters more than quality: how much control you want during the task.

19 Jun · 10 min ›
29 #ia · #ai-agents
Fable 5 vs Opus 4.8: which one to use (and the 10 tasks where the difference shows)

Fable 5 or Opus 4.8, which one should you use? A straight verdict by task type, with the 10 concrete situations where Fable gets it done and Opus 4.8 did it badly or not at all: migration at scale, code from a screenshot, long-running agents and reasoning over documents. Plus what each one costs and the cases where Opus 4.8 is still the right call.

09 Jun · 11 min ›
30 #ia · #engenharia-de-software
AI Engineer Salary in Brazil in 2026: R$ 7k to R$ 38k as an Employee, and Why Contractors Earn 50% More

There are three markets running in parallel for AI Engineers in Brazil, and each one has its own range, tax rate and negotiation criteria: CLT (salaried employment), PJ (contractor) and foreign contracts through an EOR. Here are the three 2026 ranges by level, with sources, plus the 1.5x ratio between CLT and PJ, the all-in cost a foreign employer actually pays for you, and the three specializations that move the number more than stack or tenure.

10 May · 17 min ›
31 #laravel · #php
Spec-Driven Development: A Practical Guide from PRD to Code

Vibe coding with an agent in Laravel works until the feature has business rules. Then the agent makes things up. Spec-Driven Development fixes that by turning the specification into the source of truth. In this post we walk through the PRD, spec, plan, tasks, code and tests cycle on a feature that looks silly: exporting a sales report as a PDF. PHP stack, Claude Code and Spec Kit, from scratch.

04 May · 13 min ›
Meet the Clã Beer and Code
playing