Leadership1 publisher3 min readPublished
Ben Mann's Labs group shipped Claude Code at a hit rate he puts at 20 to 30 percent, and the cheap part of that machine to copy is the two-week kill review rather than the research signal feeding it.
The Board Room · Leadership desk
Compiled by The Board RoomSomething wrong?How this is made
The persevere-or-pivot review reads as a judgment on ideas, and its more demanding function is logistical. When a bet is killed, the people on it move to another task, and dead prototypes are sometimes merged into surviving ones: the Claude Chrome extension picked up work from ideas that did not stand on their own [15]. At two-week intervals that is roughly 26 decision points a year [18], a cadence that only holds if there is somewhere for freed engineers to land and no career cost in landing there [4].
Mann puts the hit rate at 20 to 30 percent [5], which means 70 to 80 percent of bets are ended or absorbed [17]. He also says the team has had more hits than he expected, and that the hits went far [16]. The record lacks a denominator: a rate quoted without a count of bets cannot be turned into a cost per hit or an expected value, so the honest reading is that the pattern produced at least three shipped things from about 20 people, not that its economics have been demonstrated [21].
The obvious comparison is to Area 120 with better models. Mann does not dispute the lineage: he cites Bell Labs as a model, worked on Area 120 at Google in 2018, and Google's X helped build Waymo and Google Brain [9]. What separates Labs is early sight of research, not the structure: Mann's developers hear what is going well inside Anthropic, and heard that new models were showing strength at agentic coding well before Claude Code was built in late 2024 [8]. Labs was created in 2024 [2] and the preview shipped in February 2025 [7], so the team's largest hit landed within roughly a year of the team existing [19]. That schedule owes more to knowing what the models could suddenly do than to any ideation ritual.
The board-deck version of this is a 20-person unit with two-week gates, and that deck hides the authority underneath it. Mann first had to convince Anthropic that it should release products at all rather than ride philanthropic funding toward advanced AI [10], and Labs now pushes other teams' researchers toward capabilities it wants, including better audio understanding of Amharic and the visual outputs behind Claude Design [20]. It also told a new engineer his proposed code analysis tool was not a big enough idea and gave him a larger brief instead [6]. Two-week gates cost nothing; the scarce resource is a sponsor who can redirect research and overrule a new hire's scope.
The near-term pressure comes from the listing. Anthropic is preparing to go public, with Labs expected to keep converting frontier models into products [11], and a write-off rate that internal sponsors treat as the price of the method looks different in a shareholder report. This quarter's decision for anyone borrowing the pattern is whether the sponsor can absorb that ratio in public view; the test over the next several years is whether Labs graduates a second product at Claude Code's scale [3].
Ranked by verification strength, evidence, and original report placement.
Anthropic's Labs team is a rotating cast of around 20 employees tasked with spinning up entirely new ideas, led by Anthropic cofounder Ben Mann, and counts leaders including Instagram cofounder Mike Krieger.
Mann created Labs in 2024 to freely ideate on new product ideas.
Labs built Claude Code, the Model Context Protocol that connects AI agents to data sources, and this year's Claude Design; good ideas graduate out of Labs into their own teams within the company.
Labs works in cycles, bringing prototypes, or bets, up for a persevere or pivot review about every two weeks; bad ideas are tossed or merged into other ideas and their employees jump to another task.
Mann said Labs has about a 20% to 30% success rate with its ideas.
Mann assigned the coding task to then-new employee Boris Cherny, who had originally suggested making a code analysis tool; Mann told him it was not a big enough idea, and said new employees need to raise their level of ambition much higher.
Follow any of these and your For You feed starts watching them — no settings page required.
Evidence-backed comparisons of source perspectives and observed adoption signals. Read the methodology
Which Builder, Operator, and Investor concerns the observed source mix emphasized—not a truth score.
Evidence, demonstrated adoption, hype gap, incentives, and confidence are assessed independently, each on its own current evidence. How these are measured.
One interview, one outside voice
Every mechanism and every number in our coverage comes from a single Business Insider conversation with Ben Mann: the hit rate, the fortnightly review, the Amharic audio push, the story of Cherny's rejected proposal. The parts anyone can check from outside are the products themselves, and the one external voice, Forrester's Gualtieri, speaks to competition rather than to how Labs performs.
Products shipped, usage unverified
The team's products are real and out: Claude Code went from internal prototype to preview to its own team under Cherny, the Model Context Protocol is in the field connecting agents to data sources, Claude Design shipped this year, and Mann says other Anthropic engineering teams have copied the bets format. Against that, Business Insider reports no users, no revenue and no install figures for any of them, so the popularity rests on the reporter's characterisation.
Credit assigned faster than it is measured
Business Insider opens by giving a 20-person group a lot of credit for Anthropic's rise and calls Claude Code the definitive AI tool of the year. The scorecard offered for that is Mann's round 20 to 30 percent with the denominator missing, and by the piece's own telling it was the model releases across the following year that turned a February 2025 preview into the popular product. Our coverage finds the process description well supported but the credit assigned to the team unverified.
A cofounder grading his own team before a listing
Mann is both an Anthropic cofounder and the founder of the group being assessed, and he is describing its record while the company prepares to go public, which is also the frame Business Insider chooses. The counterweight comes from a Forrester analyst whose firm sells guidance to the enterprises now deciding whether Claude tools displace what they buy from Adobe and Figma.
Firm on how it runs, soft on how well
The procedural claims are specific enough to trust: the cadence, the four-person graduation threshold, who was assigned what and when, and a preview date that is a matter of record. The performance claim is a founder's round number with no bets counted and no second account, so our coverage can speak with confidence about how Labs operates but cannot verify whether 20 to 30 percent is actually a good result.
build
Every one of 364 skills in the biggest Claude Code repository leaves auto-invocation on1 publisher
invest
Box's billings ran eight points ahead of revenue. The AI bill was 20 basis points.1 publisher
product
Adobe puts more than 70 creative tools behind Slackbot on Slack's two top tiers2 publishers
build
Anthropic gives Claude Code a live window into a running iOS app1 publisher
Publishers with included, body-backed reporting in this cluster.
1 article · September 7, 2026