-
MotionLLM: Understanding Human Behaviors from Human Motions and Videos
Paper • 2405.20340 • Published • 20 -
Spectrally Pruned Gaussian Fields with Neural Compensation
Paper • 2405.00676 • Published • 10 -
Paint by Inpaint: Learning to Add Image Objects by Removing Them First
Paper • 2404.18212 • Published • 29 -
LoRA Land: 310 Fine-tuned LLMs that Rival GPT-4, A Technical Report
Paper • 2405.00732 • Published • 122
Collections
Discover the best community collections!
Collections including paper arxiv:2510.18212
-
ATLAS: Adaptive Transfer Scaling Laws for Multilingual Pretraining, Finetuning, and Decoding the Curse of Multilinguality
Paper • 2510.22037 • Published • 19 -
Less is More: Recursive Reasoning with Tiny Networks
Paper • 2510.04871 • Published • 501 -
The Dragon Hatchling: The Missing Link between the Transformer and Models of the Brain
Paper • 2509.26507 • Published • 538 -
Scaling Language-Centric Omnimodal Representation Learning
Paper • 2510.11693 • Published • 100
-
A Definition of AGI
Paper • 2510.18212 • Published • 35 -
Nemotron Elastic: Towards Efficient Many-in-One Reasoning LLMs
Paper • 2511.16664 • Published • 26 -
Step-Audio-R1 Technical Report
Paper • 2511.15848 • Published • 53 -
Unveiling Intrinsic Dimension of Texts: from Academic Abstract to Creative Story
Paper • 2511.15210 • Published • 89
-
Less LLM, More Documents: Searching for Improved RAG
Paper • 2510.02657 • Published • 2 -
ExGRPO: Learning to Reason from Experience
Paper • 2510.02245 • Published • 80 -
A Definition of AGI
Paper • 2510.18212 • Published • 35 -
Lumine: An Open Recipe for Building Generalist Agents in 3D Open Worlds
Paper • 2511.08892 • Published • 201
-
HoloScene: Simulation-Ready Interactive 3D Worlds from a Single Video
Paper • 2510.05560 • Published • 7 -
TaTToo: Tool-Grounded Thinking PRM for Test-Time Scaling in Tabular Reasoning
Paper • 2510.06217 • Published • 63 -
Less is More: Recursive Reasoning with Tiny Networks
Paper • 2510.04871 • Published • 501 -
Fast-dLLM v2: Efficient Block-Diffusion LLM
Paper • 2509.26328 • Published • 55
-
Emu3.5: Native Multimodal Models are World Learners
Paper • 2510.26583 • Published • 108 -
RECALL: REpresentation-aligned Catastrophic-forgetting ALLeviation via Hierarchical Model Merging
Paper • 2510.20479 • Published • 11 -
A Definition of AGI
Paper • 2510.18212 • Published • 35 -
Video-As-Prompt: Unified Semantic Control for Video Generation
Paper • 2510.20888 • Published • 45
-
Less is More: Recursive Reasoning with Tiny Networks
Paper • 2510.04871 • Published • 501 -
SpikingBrain Technical Report: Spiking Brain-inspired Large Models
Paper • 2509.05276 • Published • 4 -
Self-Adapting Language Models
Paper • 2506.10943 • Published • 7 -
The Art of Scaling Reinforcement Learning Compute for LLMs
Paper • 2510.13786 • Published • 31
-
Describe What You See with Multimodal Large Language Models to Enhance Video Recommendations
Paper • 2508.09789 • Published • 5 -
MM-BrowseComp: A Comprehensive Benchmark for Multimodal Browsing Agents
Paper • 2508.13186 • Published • 19 -
ZARA: Zero-shot Motion Time-Series Analysis via Knowledge and Retrieval Driven LLM Agents
Paper • 2508.04038 • Published • 1 -
Prompt Orchestration Markup Language
Paper • 2508.13948 • Published • 48
-
MotionLLM: Understanding Human Behaviors from Human Motions and Videos
Paper • 2405.20340 • Published • 20 -
Spectrally Pruned Gaussian Fields with Neural Compensation
Paper • 2405.00676 • Published • 10 -
Paint by Inpaint: Learning to Add Image Objects by Removing Them First
Paper • 2404.18212 • Published • 29 -
LoRA Land: 310 Fine-tuned LLMs that Rival GPT-4, A Technical Report
Paper • 2405.00732 • Published • 122
-
Emu3.5: Native Multimodal Models are World Learners
Paper • 2510.26583 • Published • 108 -
RECALL: REpresentation-aligned Catastrophic-forgetting ALLeviation via Hierarchical Model Merging
Paper • 2510.20479 • Published • 11 -
A Definition of AGI
Paper • 2510.18212 • Published • 35 -
Video-As-Prompt: Unified Semantic Control for Video Generation
Paper • 2510.20888 • Published • 45
-
ATLAS: Adaptive Transfer Scaling Laws for Multilingual Pretraining, Finetuning, and Decoding the Curse of Multilinguality
Paper • 2510.22037 • Published • 19 -
Less is More: Recursive Reasoning with Tiny Networks
Paper • 2510.04871 • Published • 501 -
The Dragon Hatchling: The Missing Link between the Transformer and Models of the Brain
Paper • 2509.26507 • Published • 538 -
Scaling Language-Centric Omnimodal Representation Learning
Paper • 2510.11693 • Published • 100
-
A Definition of AGI
Paper • 2510.18212 • Published • 35 -
Nemotron Elastic: Towards Efficient Many-in-One Reasoning LLMs
Paper • 2511.16664 • Published • 26 -
Step-Audio-R1 Technical Report
Paper • 2511.15848 • Published • 53 -
Unveiling Intrinsic Dimension of Texts: from Academic Abstract to Creative Story
Paper • 2511.15210 • Published • 89
-
Less LLM, More Documents: Searching for Improved RAG
Paper • 2510.02657 • Published • 2 -
ExGRPO: Learning to Reason from Experience
Paper • 2510.02245 • Published • 80 -
A Definition of AGI
Paper • 2510.18212 • Published • 35 -
Lumine: An Open Recipe for Building Generalist Agents in 3D Open Worlds
Paper • 2511.08892 • Published • 201
-
Less is More: Recursive Reasoning with Tiny Networks
Paper • 2510.04871 • Published • 501 -
SpikingBrain Technical Report: Spiking Brain-inspired Large Models
Paper • 2509.05276 • Published • 4 -
Self-Adapting Language Models
Paper • 2506.10943 • Published • 7 -
The Art of Scaling Reinforcement Learning Compute for LLMs
Paper • 2510.13786 • Published • 31
-
HoloScene: Simulation-Ready Interactive 3D Worlds from a Single Video
Paper • 2510.05560 • Published • 7 -
TaTToo: Tool-Grounded Thinking PRM for Test-Time Scaling in Tabular Reasoning
Paper • 2510.06217 • Published • 63 -
Less is More: Recursive Reasoning with Tiny Networks
Paper • 2510.04871 • Published • 501 -
Fast-dLLM v2: Efficient Block-Diffusion LLM
Paper • 2509.26328 • Published • 55
-
Describe What You See with Multimodal Large Language Models to Enhance Video Recommendations
Paper • 2508.09789 • Published • 5 -
MM-BrowseComp: A Comprehensive Benchmark for Multimodal Browsing Agents
Paper • 2508.13186 • Published • 19 -
ZARA: Zero-shot Motion Time-Series Analysis via Knowledge and Retrieval Driven LLM Agents
Paper • 2508.04038 • Published • 1 -
Prompt Orchestration Markup Language
Paper • 2508.13948 • Published • 48