#pretraining
- Video Generation Models are General-Purpose Vision Learners Google DeepMind research
- ByteDance Seed paper studies how high-quality domain data should repeat when scaling LLMs ByteDance research
- Scalable Visual Pretraining for Language Intelligence research
- Understanding Reasoning from Pretraining to Post-Training research
- Towards Physics of Multimodal Pretraining: Knowledge Flow, Modality Synergy, Early Unification, and Recipes Meta AI research