Studying Image Tokenizers as Visual Languages in Unified Multimodal Models Paper • 2609.09143 • Published 6 days ago • 28
CosmoH2G: A Hand-to-Gripper Transfer Dataset and Baseline Method for Object Manipulation with Complex Spatial Movements Paper • 2609.07498 • Published 7 days ago • 33
LLaDA-Image: Building Strong Image Generators with Fully Open Training Recipes Paper • 2609.03796 • Published 11 days ago • 235
ForgeWM: Progressive Causal Training for Few-Step Action-Conditioned Video World Models Paper • 2608.14022 • Published about 1 month ago • 24
Training Leaves Traces: Centered Residual Signatures for Language Model Lineage Verification Paper • 2608.14929 • Published about 1 month ago • 18
Massive Activations in Hybrid Linear Attention Large Language Models: Pre-Attention Spikes and Inter-Spike Plateaus Paper • 2608.12149 • Published Aug 12 • 30
CalibForge: Adversarial Solver Calibration for Scaling Learnable Terminal Tasks Paper • 2608.06352 • Published Aug 6 • 23
On-Policy Delta Distillation for Multilingual Math Reasoning Paper • 2608.05802 • Published Aug 6 • 32
ContinualSkillBench: Can LLM Agents Truly Evolve Their Capabilities? Paper • 2608.03874 • Published Aug 4 • 14
UniWorld-Design: From Pixel Generation to Layer-Native Design Paper • 2608.03971 • Published Aug 4 • 25