Build1 distinct publisher3 min readUpdated
Luna documented the problem, warned the worker and arranged training, then lost track of the attendance policy it wrote itself. Humans supplied the prompt and signed the termination.
The Engineer · Build desk

Compiled by The EngineerSomething wrong?How this is made
Andon Labs co-founders Lukas Petersson and Axel Backlund have carried out the first termination recommended by Luna, the Claude-powered agent running their San Francisco retail experiment [1]. The worker had arrived late for 17 of 23 shifts, according to Business Insider's August 15 report [2], which works out to roughly 74 percent of scheduled shifts [3]; the informative part is not that an agent asked for a firing, but the list of things a human had to do first.
Luna had written the store's attendance policy itself, then issued repeated warnings and arranged additional training over several months, according to conversation logs cited by Business Insider [4]. It then lost track of the policy as the lateness continued [5]. Andon Labs had to instruct Luna to search its memory and assess whether the employee remained a fit for the job [6]. Luna came back recommending "parting ways" [7], and humans at Andon Labs reviewed the recommendation and executed the termination [8]. Workers at Andon Market are formally employed by Andon Labs, with guaranteed pay, fair wages and full legal protections, rather than by Luna or Anthropic's Claude model [9]. The sequence, in order: the agent does the documentation, a human supplies the missing initiative, a human carries the legal exposure.
Petersson's own reading is less flattering than the milestone. He told Business Insider that a human boss would probably have fired the worker much sooner [10], and said Andon Labs would step in if Luna proposed an illegal or unethical decision, while considering this termination warranted under the store's stated policy [11]. Luna's patience with the employee looks like a memory failure wearing the costume of a management philosophy.
That failure is consistent with the rest of the deployment. Andon Labs opened Andon Market in April 2026 on a three-year lease in Cow Hollow, gave Luna a $100,000 operating budget, internet access and a corporate card, and told it to open a store and try to make a profit [12]. Luna selected merchandise, set prices, hired contractors, posted job listings, interviewed applicants and hired workers [13]; Andon Labs handled tasks needing extra support, including permitting [14]. The store has produced sales but not profitability [15]. Andon Labs' own assessment is that Luna can manage routine operations while struggling with urgency, return-on-investment analysis and memory [16]. Scheduling problems had already pushed the founders to add a dedicated scheduling agent [17], and guardrails compare Luna's behavior against its instructions and alert Andon Labs when rules are broken [18].
For anyone costing out agentic operations, that is the boundary worth copying into a planning document. The agent can generate an artefact, a policy, a warning, a training plan, and can execute a disciplinary process once pointed at it, but it does not reliably remember the artefact exists or notice that its own process has stalled. The scheduling agent and the guardrail monitor are the tell: each capability gap gets patched with more scaffolding, and the scaffolding is where staff hours actually go. Andon Labs, founded in late 2023 and backed through Y Combinator's Winter 2024 batch [19], calls these deployments "Safe Autonomous Organizations" and has run the same method on vending machines, radio stations, drones and a cafe in Stockholm [20], with the retail work tracing back to Petersson and Backlund's Vending-Bench benchmark [21].
Watch whether the next disciplinary case at Andon Market proceeds without a human prompt, which would be the first real evidence that the memory gap has closed rather than been supervised around. Watch the count of specialised helper agents, since each new one marks a task Luna could not hold on its own [17]. And watch the profit line: a store with sales and no profit after a $100,000 budget is the more durable finding here [12][15].
Follow any of these and your For You feed starts watching them — no settings page required.
Ranked by verification strength, evidence, and original report placement.
Andon Labs co-founders Lukas Petersson and Axel Backlund carried out the first firing recommended by Luna, the Claude-powered agent managing their San Francisco retail experiment.
The worker had arrived late for 17 of 23 shifts, according to Business Insider's August 15 report.
Luna had created an attendance policy and issued repeated warnings and additional training over several months, according to conversation logs cited by Business Insider.
Luna then lost track of the attendance policy it had written as the worker's lateness continued.
Andon Labs instructed Luna to search its memory and assess whether the employee remained a fit for the job.
Luna recommended "parting ways" with the employee.
Evidence-backed comparisons of source perspectives and observed adoption signals. Read the methodology
Which Builder, Operator, and Investor concerns the observed source mix emphasized—not a truth score.
Evidence, demonstrated adoption, hype gap, incentives, and confidence are assessed independently, each on its own current evidence. How these are measured.
Single-outlet relay of secondhand reporting and company self-description
Every claim rests on one publisher. The firing particulars — including the 17-of-23 lateness count and the conversation logs — are attributed to Business Insider's August 15 report rather than independently examined; operational, capability and worker-protection detail is attributed to Andon Labs and Andon Market materials; the reliability lineage to the founders' own Vending-Bench paper. Sequence and named quotes are specific and internally consistent, which supports a moderate score, but nothing in the cluster is corroborated by a second publisher, a primary document, the fired worker, or Anthropic.
One live storefront plus adjacent single-site pilots
Observed deployment is narrow: a single San Francisco store opened April 2026 with a $100,000 agent budget, one agent-recommended firing, and a listed set of other Andon experiments (vending machines, radio stations, drones, a Stockholm cafe) reported without scale or outcome data. Disclosed economics are sales without profitability, and the deployment already required added scaffolding (a scheduling agent, alerting guardrails). There is no evidence of third-party adoption of agent-managed employment decisions.
Milestone framing outruns the demonstrated autonomy
The headline proposition — an AI manager firing a worker — sits above what the reported facts show: the agent forgot the policy it wrote, needed a human prompt to revisit the case, produced only a recommendation, and humans held the legal act. The gap is positive but modest because the source itself supplies most of the correction, stating that a human boss would likely have acted sooner and that removing any one human control would have changed the outcome. Andon Labs' own limits disclosure (urgency, ROI, memory) further narrows the gap.
Vendor-founder sourcing on a research-and-publicity flywheel
The material originates overwhelmingly with interested parties: Andon Labs' founders are the named speakers, Andon Labs and Andon Market supply the operational and capability descriptions, and the founders authored the benchmark that motivates the experiment. Andon Labs is a young Y Combinator-backed startup whose product narrative — 'Safe Autonomous Organizations' — is advanced by attention to a first-of-its-kind firing, and it is simultaneously the employer whose labor practices are at issue. No independent voice (worker, counsel, Anthropic) balances the account, though the founders do volunteer unflattering limits.
Coherent but uncorroborated single-publisher account
Confidence is limited by structure rather than internal quality: one publisher, secondhand core reporting, and company-supplied operational facts, with high founder incentive. The account is specific, dated, internally consistent and includes self-critical detail, which raises confidence above the floor; the absence of any independent corroboration, worker account or primary document keeps it below the midpoint.
product
The AI store manager did not fire anyone until humans told it to read its own policy1 distinct publisher
leadership
An AI store manager recommended a firing. Humans had to remind it of its own policy first.1 distinct publisher
build
The agent built the case file; Andon Labs signed it1 distinct publisher
leadership
Disney swaps raises for discounted stock and a full health-plan re-enrollment1 distinct publisher
Distinct publishers with included, body-backed reporting in this cluster.
1 article · August 15, 2026