OpenART: Scaling Agent Red Teaming via Open-Ended Environment Evolution Paper • 2608.00677 • Published 16 days ago • 253
Don't Scroll Back: Missing-Evidence Memory for Streaming Dialogue Summarization Paper • 2608.09043 • Published 7 days ago • 7
MASS: Multiplayer World Models with Authoritative Shared State Paper • 2608.06257 • Published 7 days ago • 16
DataSpace: Benchmarking Data Agents for Verifiable Analytics over Heterogeneous Workspaces Paper • 2608.03451 • Published 13 days ago • 33
UniWorld-Design: From Pixel Generation to Layer-Native Design Paper • 2608.03971 • Published 12 days ago • 21
RecHarness: A Bandit-Routed Agentic Harness for Self-Evolving Recommender Systems Paper • 2607.29241 • Published 17 days ago • 11
QQWorld: Quantile-Quantile Matching for World Model Regularization Paper • 2607.28415 • Published 18 days ago • 30