Collections
Discover the best community collections!
Collections including paper arxiv:2603.19312
-
General agents need world models
Paper • 2506.01622 • Published • 2 -
Robust agents learn causal world models
Paper • 2402.10877 • Published • 3 -
Understanding World or Predicting Future? A Comprehensive Survey of World Models
Paper • 2411.14499 • Published • 1 -
LeWorldModel: Stable End-to-End Joint-Embedding Predictive Architecture from Pixels
Paper • 2603.19312 • Published • 50
-
UCFE: A User-Centric Financial Expertise Benchmark for Large Language Models
Paper • 2410.14059 • Published • 63 -
Sketch-of-Thought: Efficient LLM Reasoning with Adaptive Cognitive-Inspired Sketching
Paper • 2503.05179 • Published • 46 -
Token-Efficient Long Video Understanding for Multimodal LLMs
Paper • 2503.04130 • Published • 97 -
GoT: Unleashing Reasoning Capability of Multimodal Large Language Model for Visual Generation and Editing
Paper • 2503.10639 • Published • 53
-
Grandmaster-Level Chess Without Search
Paper • 2402.04494 • Published • 70 -
Can Mamba Learn How to Learn? A Comparative Study on In-Context Learning Tasks
Paper • 2402.04248 • Published • 32 -
Self-Play Preference Optimization for Language Model Alignment
Paper • 2405.00675 • Published • 29 -
Direct Nash Optimization: Teaching Language Models to Self-Improve with General Preferences
Paper • 2404.03715 • Published • 62
-
A guide to convolution arithmetic for deep learning
Paper • 1603.07285 • Published • 1 -
DeepSeek-Prover: Advancing Theorem Proving in LLMs through Large-Scale Synthetic Data
Paper • 2405.14333 • Published • 48 -
FlashMemory-DeepSeek-V4: Lightning Index Ultra-Long Context via Lookahead Sparse Attention
Paper • 2606.09079 • Published • 68 -
LeWorldModel: Stable End-to-End Joint-Embedding Predictive Architecture from Pixels
Paper • 2603.19312 • Published • 50
-
General agents need world models
Paper • 2506.01622 • Published • 2 -
Robust agents learn causal world models
Paper • 2402.10877 • Published • 3 -
Understanding World or Predicting Future? A Comprehensive Survey of World Models
Paper • 2411.14499 • Published • 1 -
LeWorldModel: Stable End-to-End Joint-Embedding Predictive Architecture from Pixels
Paper • 2603.19312 • Published • 50
-
UCFE: A User-Centric Financial Expertise Benchmark for Large Language Models
Paper • 2410.14059 • Published • 63 -
Sketch-of-Thought: Efficient LLM Reasoning with Adaptive Cognitive-Inspired Sketching
Paper • 2503.05179 • Published • 46 -
Token-Efficient Long Video Understanding for Multimodal LLMs
Paper • 2503.04130 • Published • 97 -
GoT: Unleashing Reasoning Capability of Multimodal Large Language Model for Visual Generation and Editing
Paper • 2503.10639 • Published • 53
-
A guide to convolution arithmetic for deep learning
Paper • 1603.07285 • Published • 1 -
DeepSeek-Prover: Advancing Theorem Proving in LLMs through Large-Scale Synthetic Data
Paper • 2405.14333 • Published • 48 -
FlashMemory-DeepSeek-V4: Lightning Index Ultra-Long Context via Lookahead Sparse Attention
Paper • 2606.09079 • Published • 68 -
LeWorldModel: Stable End-to-End Joint-Embedding Predictive Architecture from Pixels
Paper • 2603.19312 • Published • 50
-
Grandmaster-Level Chess Without Search
Paper • 2402.04494 • Published • 70 -
Can Mamba Learn How to Learn? A Comparative Study on In-Context Learning Tasks
Paper • 2402.04248 • Published • 32 -
Self-Play Preference Optimization for Language Model Alignment
Paper • 2405.00675 • Published • 29 -
Direct Nash Optimization: Teaching Language Models to Self-Improve with General Preferences
Paper • 2404.03715 • Published • 62