Skip to content

BuildNot yet confirmed elsewhere1 publisher2 min readPublished

Prime Intellect's 2,000-agent Rust port of Prime Agent is benchmarked only on harness overhead

Prime Intellect says over 2,000 agents rewrote Prime Agent in Rust, using more than 10,000 sandboxes and 200 billion-plus tokens. Its speed and memory gains were measured with inference excluded, so they describe harness overhead and leave coding-task time untested.

The Engineer · Build desk

How we use AISend a correction

Illustration accompanying Prime Intellect's 2,000-agent Rust port of Prime Agent is benchmarked only on harness overhead
Generated illustration

What happened

  • A root agent split the port into dependency-ordered tasks and coordinated separate planning, implementation, review and verification agents, so no agent approved its own changes.
  • Parity checks ran the Rust and TypeScript builds through scripted tasks and compared terminal frames, session transcripts, provider requests and daemon-protocol messages.
  • On fresh four-core, 8 GB sandboxes, cold-start input-ready latency fell from 736.1 milliseconds for TypeScript to 51.9 milliseconds for Rust.
  • Process-tree memory on a 10 MiB session fell from 1,130.0 MB under TypeScript to 237.3 MB under Rust, according to the company's hillclimb table.

Why it matters

  • constraint The 684 ms saving comes once per session start, so whether it shortens a customer's coding task depends on the inference time the benchmark left out.
  • capability Dropping a large session from about 14 percent to about 3 percent of an 8 GB sandbox leaves room for more persistent sessions and subagents on the same machine.
  • decision Teams pricing an agent-driven port have to count weeks of human follow-up alongside tokens and sandboxes, since passing parity tests did not make the Rust build ready to ship.

Both speed figures come from a scripted model, with inference excluded [13]. They measure the harness's startup and runtime overhead. On Prime Intellect's test, Rust's cold start is about 14 times faster than TypeScript's [18]. For that gain to reach a customer, harness startup would have to be a sizable share of a session's wall clock. The test removed the model time that decides that share. RuntimeWire concluded the benchmarks do not establish better end-to-end results for customers [16].

The memory figure has a better chance of transferring. Large-session process-tree RSS fell by a factor of about 4.8 [19]. Prime Agent keeps a persistent daemon that streams model output, runs tools and coordinates agents across sessions. Prime Intellect argues Rust's native execution and compile-time checks suit that long-running workload [15]. We'd expect the saving to count most where many sessions share one sandbox. The rewrite also adds Windows support, the company says [3]. Comparisons with other harnesses are weaker ground. The company's own post says there is no common benchmark standard and those results should be treated cautiously [14].

The process behind the port is the better engineering in the release. The root agent wrote no product code, so it could spend its time on task allocation and integration [6]. Anyone who has worked under a good engineering manager will recognise the job description. Verification used the TypeScript build as the oracle. For a port we think that is the right choice, because the old program is the specification and any divergence shows up as a diff against it [7]. Agents also audited features component by component [8].

The scripts still missed things. Prime Intellect moved its own agents and staff onto the Rust build, found gaps the parity checks had not caught, and spent the following weeks fixing bugs, finishing features and polishing the interface [10]. Engineers directed that follow-up and reviewed the results [9]. The agent run itself took more than two weeks [1].

The scale is large. At the stated floor figures, the run averaged about 100 million tokens per agent [21]. The rewrite's tokens, all from Prime Inference's GLM-5.3 endpoint [2], equal about 2.5 percent of everything Prime Agent has processed since its August launch, by the company's tally [22]. Prime Intellect does not break out active users, retention or the commercial share of that usage [17].

What to watch

  • End-to-end task timings with live inference for the Rust and TypeScript builds, from Prime Intellect or an outside tester.
  • A shared benchmark for comparing agent harnesses; Prime Intellect says none exists yet.
  • Bug reports from Windows users, the platform the Rust build adds.

Clarity's read

What the record supports and how the coverage leans. The claims behind it follow.

Reality

Evidence45
Adoption30
Hype gap+15
Incentives75
Confidence40
Why these scores

Claim ledger

Ranked by verification strength, evidence, and original report placement.

  1. [1]

    Prime Intellect says more than 2,000 agents rewrote Prime Agent, its open-source coding agent, in Rust over more than two weeks.

    ReportedSupportedSource: Prime Intellect, as reported by RuntimeWireView cited source
  2. [2]

    The rewrite used more than 10,000 Prime Sandboxes and over 200 billion tokens from Prime Inference's GLM-5.3 endpoint.

    ReportedSupportedSource: Prime Intellect, as reported by RuntimeWireView cited source
  3. [3]

    Published on October 9, the release says the Rust version starts faster, uses less memory and adds Windows support.

    ReportedSupportedSource: Prime Intellect, as reported by RuntimeWireView cited source

Sources

1 independent publisher whose own reporting we read for this story.

  1. runtimewire.com

    1 article · October 9, 2026

    Prime Intellect says 2,000 agents rewrote its coding agent in Rust

Share your take

Let Clarity write the post for you.

Signed-in readers get a short post drafted on this story in the register they choose — narrative, analytical, or a direct position — editable to the last word before it goes anywhere. The share buttons at the top of this story work without an account.

Topics and entities

Follow any of these and your For You feed starts watching them — no settings page required.

Topics

Entities

Loading related stories