Peter Steinberger 🦞 on X: "5.6 Terra high is underrated. Switched @clawsweeper (GitHub review bot) to it and it's ~40% faster overall with negligible quality loss. Better than 5.5 on all counts. Massively cheaper. (Tried xhigh but that negates perf wins, didn't make a noticable difference in review evals)" / Twitter
(Tried xhigh but that negates perf wins, didn't make a noticable difference in review evals)
rtk Claude Code Token Savings: A Skill Trial Benchmark
Does "rtk" reduce Claude Code token usage? Part 2 of a series where we take public “token saving” add-ons for coding agents and run the same paired A/B benchmark against each of them. Part 1 was th
Speaking to AI Agents like Cavemen Saves 65% of Tokens. We Test.
A paired A/B benchmark of the token-compression skill Caveman on Claude Code, run on SkillsBench: does it actually save tokens, and does it degrade AI agent output quality? Advertised saving: 65%.
pi-workspace/pi-workspace: Pi Workspace is a local desktop app for working with Pi across Git repositories and long-running goals. Plan, implement, and pick up where you left off.
Pi Workspace is a local desktop app for working with Pi across Git repositories and long-running goals. Plan, implement, and pick up where you left off. - pi-workspace/pi-workspace
Lessons from Building a First-Pass AI PRD Reviewer at Uber
What if every PM had a fast, context-rich first-pass reviewer before a PRD reached the review room? We built an AI PRD Evaluator at Uber to surface blind spots early, uncover prior artifacts, and give product reviews higher signal.
A Fireside Chat with Cat and Thariq from the Claude Code team
Earlier this month I hosted a fireside chat session at the AI Engineer World’s Fair with Cat Wu and Thariq Shihipar from Anthropic’s Claude Code team. We talked about Claude …