跳到正文
@rohanpaul_ai· @rohanpaul_ai · X·· 2026-08-26AI 评分57
AI 导读

Prime Agent 的技术报告发布,这一 harness 曾让 Opus 5 在 ARC-AGI-3 上达到 95.5%,报告标题为 Prime Agent: A Self-Improving RLM Harness。

正文

Prime Agent, the harness that put Opus 5 at 95.5% on ARC-AGI-3, now has a technical report.

The mechanism is a memory hierarchy. Model weights and active context sit underneath a persistent IPython session and a disk-backed store of histories, skills and prompts, and the model moves state between those levels with code instead of having it compacted away.

Long inputs stay in the REPL as variables the agent can search and transform, so long-context work becomes an information-management problem rather than a reading problem.

– arxiv. org/abs/2608.23552

Title: "Prime Agent: A Self-Improving RLM Harness"

来源:@rohanpaul_ai · x.com