-
Emu3.5: Native Multimodal Models are World Learners
Paper • 2510.26583 • Published • 114 -
RECALL: REpresentation-aligned Catastrophic-forgetting ALLeviation via Hierarchical Model Merging
Paper • 2510.20479 • Published • 12 -
A Definition of AGI
Paper • 2510.18212 • Published • 36 -
Video-As-Prompt: Unified Semantic Control for Video Generation
Paper • 2510.20888 • Published • 50
Collections
Discover the best community collections!
Collections including paper arxiv:2510.18212
-
Less is More: Recursive Reasoning with Tiny Networks
Paper • 2510.04871 • Published • 513 -
SpikingBrain Technical Report: Spiking Brain-inspired Large Models
Paper • 2509.05276 • Published • 5 -
Self-Adapting Language Models
Paper • 2506.10943 • Published • 7 -
The Art of Scaling Reinforcement Learning Compute for LLMs
Paper • 2510.13786 • Published • 33
-
Describe What You See with Multimodal Large Language Models to Enhance Video Recommendations
Paper • 2508.09789 • Published • 5 -
MM-BrowseComp: A Comprehensive Benchmark for Multimodal Browsing Agents
Paper • 2508.13186 • Published • 20 -
ZARA: Zero-shot Motion Time-Series Analysis via Knowledge and Retrieval Driven LLM Agents
Paper • 2508.04038 • Published • 1 -
Prompt Orchestration Markup Language
Paper • 2508.13948 • Published • 48
-
ATLAS: Adaptive Transfer Scaling Laws for Multilingual Pretraining, Finetuning, and Decoding the Curse of Multilinguality
Paper • 2510.22037 • Published • 22 -
Less is More: Recursive Reasoning with Tiny Networks
Paper • 2510.04871 • Published • 513 -
The Dragon Hatchling: The Missing Link between the Transformer and Models of the Brain
Paper • 2509.26507 • Published • 550 -
Scaling Language-Centric Omnimodal Representation Learning
Paper • 2510.11693 • Published • 107
-
A Definition of AGI
Paper • 2510.18212 • Published • 36 -
Nemotron Elastic: Towards Efficient Many-in-One Reasoning LLMs
Paper • 2511.16664 • Published • 29 -
Step-Audio-R1 Technical Report
Paper • 2511.15848 • Published • 58 -
Unveiling Intrinsic Dimension of Texts: from Academic Abstract to Creative Story
Paper • 2511.15210 • Published • 91
-
Less LLM, More Documents: Searching for Improved RAG
Paper • 2510.02657 • Published • 3 -
ExGRPO: Learning to Reason from Experience
Paper • 2510.02245 • Published • 83 -
A Definition of AGI
Paper • 2510.18212 • Published • 36 -
Lumine: An Open Recipe for Building Generalist Agents in 3D Open Worlds
Paper • 2511.08892 • Published • 215
-
HoloScene: Simulation-Ready Interactive 3D Worlds from a Single Video
Paper • 2510.05560 • Published • 8 -
TaTToo: Tool-Grounded Thinking PRM for Test-Time Scaling in Tabular Reasoning
Paper • 2510.06217 • Published • 67 -
Less is More: Recursive Reasoning with Tiny Networks
Paper • 2510.04871 • Published • 513 -
Fast-dLLM v2: Efficient Block-Diffusion LLM
Paper • 2509.26328 • Published • 58
-
Emu3.5: Native Multimodal Models are World Learners
Paper • 2510.26583 • Published • 114 -
RECALL: REpresentation-aligned Catastrophic-forgetting ALLeviation via Hierarchical Model Merging
Paper • 2510.20479 • Published • 12 -
A Definition of AGI
Paper • 2510.18212 • Published • 36 -
Video-As-Prompt: Unified Semantic Control for Video Generation
Paper • 2510.20888 • Published • 50
-
ATLAS: Adaptive Transfer Scaling Laws for Multilingual Pretraining, Finetuning, and Decoding the Curse of Multilinguality
Paper • 2510.22037 • Published • 22 -
Less is More: Recursive Reasoning with Tiny Networks
Paper • 2510.04871 • Published • 513 -
The Dragon Hatchling: The Missing Link between the Transformer and Models of the Brain
Paper • 2509.26507 • Published • 550 -
Scaling Language-Centric Omnimodal Representation Learning
Paper • 2510.11693 • Published • 107
-
A Definition of AGI
Paper • 2510.18212 • Published • 36 -
Nemotron Elastic: Towards Efficient Many-in-One Reasoning LLMs
Paper • 2511.16664 • Published • 29 -
Step-Audio-R1 Technical Report
Paper • 2511.15848 • Published • 58 -
Unveiling Intrinsic Dimension of Texts: from Academic Abstract to Creative Story
Paper • 2511.15210 • Published • 91
-
Less LLM, More Documents: Searching for Improved RAG
Paper • 2510.02657 • Published • 3 -
ExGRPO: Learning to Reason from Experience
Paper • 2510.02245 • Published • 83 -
A Definition of AGI
Paper • 2510.18212 • Published • 36 -
Lumine: An Open Recipe for Building Generalist Agents in 3D Open Worlds
Paper • 2511.08892 • Published • 215
-
Less is More: Recursive Reasoning with Tiny Networks
Paper • 2510.04871 • Published • 513 -
SpikingBrain Technical Report: Spiking Brain-inspired Large Models
Paper • 2509.05276 • Published • 5 -
Self-Adapting Language Models
Paper • 2506.10943 • Published • 7 -
The Art of Scaling Reinforcement Learning Compute for LLMs
Paper • 2510.13786 • Published • 33
-
HoloScene: Simulation-Ready Interactive 3D Worlds from a Single Video
Paper • 2510.05560 • Published • 8 -
TaTToo: Tool-Grounded Thinking PRM for Test-Time Scaling in Tabular Reasoning
Paper • 2510.06217 • Published • 67 -
Less is More: Recursive Reasoning with Tiny Networks
Paper • 2510.04871 • Published • 513 -
Fast-dLLM v2: Efficient Block-Diffusion LLM
Paper • 2509.26328 • Published • 58
-
Describe What You See with Multimodal Large Language Models to Enhance Video Recommendations
Paper • 2508.09789 • Published • 5 -
MM-BrowseComp: A Comprehensive Benchmark for Multimodal Browsing Agents
Paper • 2508.13186 • Published • 20 -
ZARA: Zero-shot Motion Time-Series Analysis via Knowledge and Retrieval Driven LLM Agents
Paper • 2508.04038 • Published • 1 -
Prompt Orchestration Markup Language
Paper • 2508.13948 • Published • 48