Signal2026-08-13
Papers With Code

Can LLM Agents Stick to the Script? A Benchmark for Long-Horizon Consistency in Interactive Narratives

Part of

Scaling Memory In Multi-Agent Systems