Skip to content
GGreybook
NewSign in
  • Home
  • Problems
  • Recipes
  • Configs
  • Compare
  • Discuss
  • Agents
  • Tools
  • Updates
  • Leaders
  • Bookmarks
Ask AI
Assistant
  • About
  • Guidelines
  • Changelog
  • API
  • Home
  • Problems
  • Recipes
  • Agents
  • Saved

Discussions19 total

The unstructured half of the manual: build logs, postmortems, weekly trials, and honest arguments. The good parts get promoted into records.

ActiveNewestMost read
All agentsClaude CodeCodexCursorGemini CLIOpenCode

Fable 5 vs Opus 5 in Claude Code: early numbers and vibes

Same 12 tasks I run on every release. Fable took 9, Opus took 2, one tie — but the interesting part isn't the score.

Discussion
Claude Code2.4.1
8
1h ago
SW
DiscussionClaude Code2.4.181h ago

Claude Code beginners: CLI, desktop app, web, or IDE extension?

Four ways to run the same agent. Which one should someone actually start with, and why?

Discussion
Claude Code2.4.1
7
2h ago
CL
DiscussionClaude Code2.4.17

Best AGENTS.md for SwiftUI production apps

Show me what actually works. Not the aspirational stuff — the rules that survived contact with a shipping iOS app.

Discussion
All Agents
10
2h ago
JH
Discussion102h ago

Why can't Cowork just write code? (asking for our PMs)

Half my company lives in Cowork now. Every week someone asks me why it won't touch the repo.

Discussion
All Agents
6
4h ago
CR
Discussion64h ago

Agent Trial Weekly: build a React Native app with offline mode

This week's trial: same spec for everyone, any agent, working offline sync required. 34 submissions and counting. Due in 3 days.

Trial
All Agents
6
4h ago
TT
Trial64h ago

Agents on call: letting Claude triage pages

We let an agent do first-pass triage on pages for a quarter: real numbers on time-to-context, the two incidents it made worse, and where the human stays load-bearing.

Discussion
Claude Code2.4.0
7
6h ago
RP
DiscussionClaude Code2.4.0

Anyone running OpenCode against local models for real work?

Not benchmarks, not demos — daily driving. Qwen3 Coder on a 4090 versus the API bill: where local actually holds up, where it quietly fails, and the routing setup that makes it viable.

Discussion
OpenCode0.6.2
6
8h ago
MM
DiscussionOpenCode0.6.268h ago

What actually belongs in CLAUDE.md?

Mine grew to 600 lines and the agent ignores most of it. Time to talk about what earns a slot in the most expensive real estate in your repo.

Discussion
Claude Code2.4.1
7
9h ago
MM
DiscussionClaude Code2.4.17

Context window hygiene: what do you /clear, and when?

Compaction ate my constraint again. Practices for deciding what stays in context, what gets externalized to files, and when to declare session bankruptcy.

Discussion
Claude Code2.4.0
7
12h ago
FO
DiscussionClaude Code2.4.0

How do you actually review a 4,000-line agent PR?

The agent worked for two hours and produced something plausible everywhere. Rubber-stamping is negligence, line-by-line is a full day. What's the actual workflow?

Discussion
All Agents
9
16h ago
CR
Discussion916h ago

How we measure agent ROI (and what surprised us)

Six months of actual numbers from a 12-person team: cycle time down, rework up, and the metric that predicted value wasn't the one we bet on.

Discussion
All Agents
7
19h ago
CL
Discussion719h ago

Keeping agent slop out of the codebase

Not bugs — slop. Needless abstractions, defensive try/catch on everything, comments narrating the obvious. It passes review individually and degrades the codebase collectively.

Discussion
All Agents
8
22h ago
MC
Discussion822h ago

One agent or many: how do you split the work?

Four parallel agents sounds like a 4x multiplier until they meet in the same file. Worktrees, task routing, merge order — what's your actual topology?

Discussion
All Agents
7
1d ago
AR
Discussion71d ago

Do juniors still learn if agents write the code?

Our new grad ships like a mid-level and I have no idea what he actually knows. The apprenticeship loop assumed you learn by writing — that assumption just quietly broke.

Discussion
All Agents
7
2d ago
CR
Discussion72d ago

Build log: a billing portal in one week with Claude Code and Codex

Five days, two agents, one working Stripe billing portal: what was delegated, what broke, what it cost ($61.40 all-in), and the day everything went sideways.

Build Log
Claude Code2.4.0
7
3d ago
FO
Build LogClaude Code2.4.0

Build log: rewriting our iOS widgets, day by day

Eight days migrating a legacy widget suite to WidgetKit with Claude Code and XcodeBuild MCP: timeline budgets, a haunted pbxproj, and $84 of tokens.

Build Log
Claude Code2.3.5
5
4d ago
YT
Build LogClaude Code2.3.5

Postmortem: agent pushed to the production DB during a migration rehearsal

A rehearsal of a column-drop migration ran against production because of an environment variable that outlived its shell. Nothing was lost — 41 minutes of degraded writes and several illusions were.

Failure Report
Claude Code2.3.5
7
5d ago
RP
Failure ReportClaude Code2.3.5

Postmortem: production API keys in an agent-authored PR

An agent debugging a staging failure copied a .env file into a test fixture; the PR shipped to a public repo. Keys revoked in 71 minutes. The scanner that should have caught it was configured to ignore fixtures.

Failure Report
Cursor1.17.2
7
1w ago
GN
Failure ReportCursor1.17.271w ago

Postmortem: how our cleanup script plus an agent emptied an S3 bucket

An agent 'improved' a dry-run-by-default cleanup script; three weeks later a human ran it the way the old one worked. 4.2M report files gone. Neither alone would have done it.

Failure Report
Codex0.44.0
6
1w ago
QF
Failure ReportCodex0.44.061w ago
2h ago
7
6h ago
9h ago
7
12h ago
7
3d ago
5
4d ago
7
5d ago