DreamX-Creator: Democratizing Native Audio-Video Generation at 2K Resolution Paper • 2608.31106 • Published 1 day ago • 84
Lucida: Parse, Generate, and Place for Composable Real-to-Sim Scene Modeling Paper • 2608.30821 • Published 1 day ago • 53
Code as Worlds: Agentic Discovery of Executable World Representations for Physical Reasoning Paper • 2608.27549 • Published 6 days ago • 45
Magpie: Real-Time World Renderer for Interactive Games Paper • 2608.27168 • Published 6 days ago • 10
Game2World Engine: Unlocking In-the-Wild Gameplay Videos for World Model Training Paper • 2608.24680 • Published 8 days ago • 11
ReWorld: An Interactive World Model with Long-Horizon Memory Paper • 2608.23565 • Published 9 days ago • 24
PACE-Bench: Benchmarking Physics Adaptation via Code Evolution in Dynamic Environments Paper • 2608.14441 • Published 19 days ago • 28
GS-Voxel: Fitting-Free Structured Latents for Large-Scale 3DGS Generation Paper • 2608.17988 • Published 15 days ago • 4
ASI-Bench: At the Dawn of Artificial Superintelligence Paper • 2608.17271 • Published 15 days ago • 62
How Can Rhetoric Reward-Hack AI Reviewers? Dissecting Rhetorical Sensitivity in AI-Based Peer Review Paper • 2608.08975 • Published 23 days ago • 48
PlayWorld: Benchmarking World Models with Agent Players over Long-Horizon Objectives Paper • 2608.13552 • Published 20 days ago • 46
DreamX-Phi 1.0: Action-Conditioned Video World Model for Robotic Manipulation Paper • 2608.13489 • Published 20 days ago • 98
Self-Geometry: GT-Free and Plug-and-Play Test-Time Adaptation for Geometrically Consistent 3D Vision Foundation Models Paper • 2608.10708 • Published 22 days ago • 15
Co-Evolution in Agentic Systems: Toward Self-Directed Evolution Beyond Human Design Paper • 2608.10299 • Published 23 days ago • 135
MerchantBench: Benchmarking LLM Agents for Long-Term Coherence in E-Commerce Operations Paper • 2607.28956 • Published Jul 31 • 97
RL^2-VLA: Adaptive RL Latent Compositional Steering with Test-Time Scaling for Vision-Language-Action Models Paper • 2607.26991 • Published Jul 30 • 9
N_0-VTLA: Scaling Vision-Tactile-Language-Action Model with Latent Tactile Tokens Paper • 2607.23782 • Published Jul 26 • 78