Rewriting Prime Agent in Rust

(primeintellect.ai)

53 points | by piotrgrabowski 11 hours ago

15 comments

  • AbuAssar 7 hours ago
    > Overall, our Rust rewrite and following performance hillclimbing has made Prime Agent significantly faster and more resource efficient. With time to input roughly 14x faster than TypeScript and using over 80% less memory after startup

    very nice outcome of this port

  • pjmlp 1 hour ago
    There are enough compiled languages to chose from, start there.

    Naturally then there isn't source material for "we rewrote yet another slow scripting project into Go/Rust/Zig/C/C++/..." blog posts.

  • theturtletalks 10 hours ago
    So they used Prime Agent and GLM 5.3 to swarm and rewrite the code in Rust. This also shows Prime Agent doing what it preaches by rebuilding itself. Since Prime Agent is Pi under the hood, will they push a Rust rewrite to Pi? Pi extensions use Typescript so I wonder if they will work.

    I don’t see many people talk about Prime Agent, I always wondered if it could just be a Pi extension cause it seems to be a subagent orchestration agent.

    • sejje 6 hours ago
      I feel like I talk about it often enough people might think I have an agenda. For a while it felt like a superpower compared to other harnesses, although they've caught up with persistent/backgrounded agents.

      Very happy user. Loved it with deepseek-v4-flash, although I've moved on because newer models are so compelling.

      I just get good results from it; I think it's the python requirement.

      It has been very good with sub-agents, and also finding old context, for a long time. And had backgrounded agents that you can run from one instance without herdr. Herdr is kinda redundant (better UI than PA though).

      I see nobody else talk about it. It's so little talk that when I mention it on X, the devs comment on my posts sometimes.

  • mgreg 9 hours ago
    I'm interested in their Planner -> Implementer -> Reviewer -> Verifier process they used for this transition to Rust. I've see similar but curious how they actually implemented this.

    Curious how this could be applied to greenfield coding rather than just making a copy in a new language or performance optimizing.

  • helsinki 9 hours ago
    If anyone wants a 100% parity Rust port of v1.0 DeepSeek Harness, I have one here: https://github.com/trevorprater/SeekDeep-Harness
    • esafak 7 hours ago
      What benefits have you seen from using the DeepSeek Harness? What were you using before?
  • oefrha 10 hours ago
    Kinda wish people break down token usage into input (cache hit), input (cache miss) and output when talking about it. Giving a total 200B tokens number doesn’t help gauge costs.
    • ram1500natrluvr 5 hours ago
      I was checking openrouter a few days ago to see if they provided that information yet. Would be a really nice breakdown to have under total task price. Doesn't map to every task, but averages would still be nice.

      For Anthropic's reported benchmark numbers for Sonnet 5.5, when running terminal-bench 4.0, they had a comparison between $0.10 and $0.20 cache read for total task cost. That 50% cost reduction resulted in a 20.9-26.2% total cost reduction. Pretty specific though... single benchmark, model, harness, and provider. Cache reads make up ~42-52% of the total cost at the standard $0.20 price.

    • mattbusel 6 hours ago
      [flagged]
  • Tsarp 8 hours ago
    I know people tend to hate on rust rewrites. But having something that compiles to a single binary that you can just copy over and get started has its advantages especially when working with sandboxes etc.
    • binary132 7 hours ago
      Convenient yes, but also not just a rust thing!
  • esafak 7 hours ago
    Does anyone have experience to share about their Prime Agent harness? Does it do anything that regular harnesses paired with a memory plugin can't do?
  • rvz 6 hours ago
    We already knew TypeScript was the wrong language for many use cases from the start as an excuse to not learn Rust.

    Now there is no excuses to not use Rust and it just shows in raw performance alone.

    • pjmlp 1 hour ago
      Any scripting language is the wrong language, when the job is performance only delivered by compiled languages.

      Eventually AIs will generate Assembly code directly anyway, no need for intermediate 3 GL languages.

      • jasomill 1 hour ago
        That sounds like a nightmare for comprehensibility, incremental development, and maintenance.

        Combine that with the fact that, unlike other code generators, LLM output can't as a rule be reliably reproduced from the original input in the future, sounds like a recipe for unpredictable long-term costs and regular regressions.

        • pjmlp 37 minutes ago
          Spoken like an Assembly programmer when optimising compilers arrived into town.

          JIT compilers, GC workflows, PGO, and ML optimising compilers passes are also non deterministic, and yet work gets done.

          If you are curious, there is already enough work out there into this direction.

  • TokenLat 6 hours ago
    [flagged]
  • soltanov 6 hours ago
    [flagged]
  • sglim 10 hours ago
    [dead]
  • ContinuityLab 9 hours ago
    [flagged]
    • foltik 9 hours ago
      Writing comments with an LLM helps posting efficiency, but for most HN accounts, the bottleneck remains having something worth saying, not the typing.