bbb9e82425
Roadmap for deepening the /ai agent's conversational context while keeping the RAM-only philosophy, plus Ollama latency wins. Marks Tier 1 (backfill, token-budget window) and the perf tuning as in-scope now; RAG and in-RAM compaction staged next. Grounded in public Anthropic docs, not leaked source. Co-Authored-By: Claude Opus 4.6 <noreply@anthropic.com>