OpenAI 正在筹备 Codex Replay 功能,可从任意导入的对话线程测试任务执行。该功能会为每个线程单独检测历史状态与配置,默认选用 GPT-5.6 Sol、Terra、Luna 模型,支持多会话并行,选择、执行、验证与评估均在浏览器控制器中完成,而非引导式聊天流程。
OpenAI is preparing a Codex Replay feature that lets users test task execution from any imported conversation thread.
> "Start an independent Codex Replay controller on an available loopback port. In both Codex Desktop and Codex CLI, prefer an available Codex in-app browser and otherwise use your system browser."
> "Select one or more historical Claude threads, choose shared Codex models, and start their isolated implementations. Historical state and configuration are detected separately for each thread, with details available when needed."
> "View finished comparisons individually or in aggregate while remaining threads continue. Available GPT-5.6 Sol, Terra, and Luna models are selected by default."
> "Multiple replay sessions can run in parallel. Selection, execution, verification, evaluation, and results stay in the browser controller instead of a guided chat workflow."
来源:@testingcatalog · x.com