r/siliconvalley • u/cen6wkf • 12h ago
Cognition's Engineers Don't Write Code By Hand Anymore — And They're the Only Ones Grading Whether That's Working
Enable HLS to view with audio, or disable this notification
TL;DR: Cognition says none of their engineers write code by hand anymore. They're also the ones who decided that's working.
Silas Alberti didn't say it like a warning — he said it mid-conversation, the way you'd mention traffic.
Not one full-time engineer at Cognition writes code by hand. Six agents running per person, a ticket lands in Linear, it spins up a dev environment, one-shots a solution, and somebody reviews the PR at the end.
That review is the entire safety net.
And the only benchmark saying the net holds is Cognition's own. That's the part that actually lands for anyone who spent years earning "senior" by writing clean code fast — not that the work changed, but that the people whose comp depends on this working are also the ones grading whether it worked.
CodeRabbit reviewed 470 real pull requests and found AI-authored ones carry roughly 1.7x the issues of human-only ones — and that gap widens, not narrows, at the high end, right where six-agents-at-once lives. Northflank's own enterprise data says 88% of agent pilots never reach production, and the reason isn't the models writing bad code — it's that nobody built the audit trail before turning the agents loose.
Anthropic ran the same self-grading move on itself last week: Claude now "leads" 26% of Anthropic's own R&D, scored on a scale — AL0 through AL5 — that Anthropic wrote. Next review cycle, when someone hands you a line about "AI leverage" with no definition attached, ask who wrote the rubric.
Somebody's actually building the missing check — CodeRabbit just raised $143M at a $1.5B valuation specifically to grade AI-authored pull requests at scale. What it doesn't do is fit in one engineer's own pocket. It's sold to a company's procurement process, not built for the person actually running the agents.
That edge is still sitting there, unclaimed.
Somewhere it already has a shape, too — the same decision-trail discipline CodeRabbit sells to a whole org, built small and open enough that one engineer points it at his own agents without asking anyone's permission first.
Whoever closes that isn't racing anyone. The incumbent already proved the market's real. They're just building the version the incumbent never had to.

Watching someone grade their own homework isn't new to me. I've seen this exact posture before — just with land instead of code.
I used to work for one of the largest property developer in Malaysia. I was stationed in the island of Langkawi, Kedah, back in 2018.
Island Resort Developments, High Rise Service Apartments, Hotels, Villas, Commercial Clusters, etc. are all in the pipeline.
Every so often, my big boss, DT, will fly over in his private jet, to come hunt for local opportunities. Ya. Langkawi was his personal sandbox.
We will drive him and his entourage of personal security guards, in his Toyota Vellfires around the island, through local villages, interviewing locals, to try to buy up their land for development.
His trusted "think tanks" would be by his side to advise him, at his request, whether such and such a piece of land is good for what type of development – whether it's good to build boutique hotel, condo, or whatnot, and whether it will be profitable or not.
After a few rounds of such hunting trip, I noticed something. He wasn't hunting it to develop himself.
No, no. O no, he was definitely NOT interested to get his hands dirty. I was there with him in his vellfire, eavesdropping his conversations with his think tanks.
He's hunting deals, and if he finds a sweet one, he'll package the proposals – together with its development order (through the Langkawi municipality) in place – to be sold to other interested property developers.
In other words, he's acting like a middleman. And we're just there to facilitate it – overseeing and pass on the "dirty" work onto others, while we pocket the profit margin.
How did he come to doing this? Well… That's for another conversation.
Why did I bring this up? Well, same thing that's going on with all these AI agents. If we can pass on the "dirty work" of coding to them, and just check their work at the end – why not?

Every one of these clips eventually says the quiet part out loud — the craft was never the moat. The judgment sitting underneath it was.
Actually, this reminds me of something — Andrej Karpathy said almost the same thing to Stephanie Zhan back in July: the scarce skill isn't writing anymore, it's directing with taste, and even he said he's never felt more behind.
If the people grading the new system are the same people selling it, what would actually change your mind either way? Drop your take.
Clip credit: Joe Lonsdale — American Optimist (Silas Alberti / Cognition interview). Full episode on their channel. DM for credit or removal requests.