DreamX-Phi 1.0: Action-Conditioned Video World Model for Robotic Manipulation Paper • 2608.13489 • Published 7 days ago • 96
VibeWorlding: Can Multimodal Agents Construct 3D Open Worlds End-to-End? Paper • 2608.15265 • Published 5 days ago • 54
MegaParts: Scaling Part-Aware 3D Object Generation to 300 Parts via Token-Efficient Autoregressive Modeling Paper • 2608.14783 • Published 6 days ago • 17
ConceptFormer: Learning Adaptive Latent Concepts for Query-Document Alignment in Visual Document Retrieval Paper • 2608.15698 • Published 4 days ago • 4
Personalized Auto-Research: Towards a True AI Co-Scientist Paper • 2608.14881 • Published 6 days ago • 4