Rewriting Prime Agent in Rust

57 pointsposted 15 hours ago
by piotrgrabowski

17 Comments

AbuAssar

11 hours ago

> Overall, our Rust rewrite and following performance hillclimbing has made Prime Agent significantly faster and more resource efficient. With time to input roughly 14x faster than TypeScript and using over 80% less memory after startup

very nice outcome of this port

theturtletalks

13 hours ago

So they used Prime Agent and GLM 5.3 to swarm and rewrite the code in Rust. This also shows Prime Agent doing what it preaches by rebuilding itself. Since Prime Agent is Pi under the hood, will they push a Rust rewrite to Pi? Pi extensions use Typescript so I wonder if they will work.

I don’t see many people talk about Prime Agent, I always wondered if it could just be a Pi extension cause it seems to be a subagent orchestration agent.

sejje

9 hours ago

I feel like I talk about it often enough people might think I have an agenda. For a while it felt like a superpower compared to other harnesses, although they've caught up with persistent/backgrounded agents.

Very happy user. Loved it with deepseek-v4-flash, although I've moved on because newer models are so compelling.

I just get good results from it; I think it's the python requirement.

It has been very good with sub-agents, and also finding old context, for a long time. And had backgrounded agents that you can run from one instance without herdr. Herdr is kinda redundant (better UI than PA though).

I see nobody else talk about it. It's so little talk that when I mention it on X, the devs comment on my posts sometimes.

pjmlp

5 hours ago

There are enough compiled languages to chose from, start there.

Naturally then there isn't source material for "we rewrote yet another slow scripting project into Go/Rust/Zig/C/C++/..." blog posts.

mgreg

13 hours ago

I'm interested in their Planner -> Implementer -> Reviewer -> Verifier process they used for this transition to Rust. I've see similar but curious how they actually implemented this.

Curious how this could be applied to greenfield coding rather than just making a copy in a new language or performance optimizing.

oefrha

13 hours ago

Kinda wish people break down token usage into input (cache hit), input (cache miss) and output when talking about it. Giving a total 200B tokens number doesn’t help gauge costs.

ram1500natrluvr

8 hours ago

I was checking openrouter a few days ago to see if they provided that information yet. Would be a really nice breakdown to have under total task price. Doesn't map to every task, but averages would still be nice.

For Anthropic's reported benchmark numbers for Sonnet 5.5, when running terminal-bench 4.0, they had a comparison between $0.10 and $0.20 cache read for total task cost. That 50% cost reduction resulted in a 20.9-26.2% total cost reduction. Pretty specific though... single benchmark, model, harness, and provider. Cache reads make up ~42-52% of the total cost at the standard $0.20 price.

Tsarp

12 hours ago

I know people tend to hate on rust rewrites. But having something that compiles to a single binary that you can just copy over and get started has its advantages especially when working with sandboxes etc.

binary132

10 hours ago

Convenient yes, but also not just a rust thing!

esafak

10 hours ago

Does anyone have experience to share about their Prime Agent harness? Does it do anything that regular harnesses paired with a memory plugin can't do?

rvz

9 hours ago

We already knew TypeScript was the wrong language for many use cases from the start as an excuse to not learn Rust.

Now there is no excuses to not use Rust and it just shows in raw performance alone.

pjmlp

5 hours ago

Any scripting language is the wrong language, when the job is performance only delivered by compiled languages.

Eventually AIs will generate Assembly code directly anyway, no need for intermediate 3 GL languages.

jasomill

4 hours ago

That sounds like a nightmare for comprehensibility, incremental development, and maintenance.

Combine that with the fact that, unlike other code generators, LLM output can't as a rule be reliably reproduced from the original input in the future, sounds like a recipe for unpredictable long-term costs and regular regressions.

pjmlp

3 hours ago

Spoken like an Assembly programmer when optimising compilers arrived into town.

JIT compilers, GC workflows, PGO, and ML optimising compilers passes are also non deterministic, and yet work gets done.

If you are curious, there is already enough work out there into this direction.