• Home
  • Publication
  • Experience
  • Selected Publications
    • Scaling Native Multimodal Pre-Training From Scratch
    • LLaDA2.0-Uni: Unifying Multimodal Understanding and Generation with Diffusion Large Language Model
    • Diversity or Precision? A Deep Dive into Next Token Prediction
    • One-Token Rollout: Guiding Supervised Fine-Tuning of LLMs with Policy Gradient
    • Reinforcement Learning on Pre-Training Data
    • Grove MoE: Towards Efficient and Superior MoE LLMs with Adjugate Experts
    • On-Policy Optimization with Group Equivalent Preference for Multi-Programming Language Understanding
    • ToTRL: Unlock LLM Tree-of-Thoughts Reasoning Potential through Puzzles Solving
    • Efficient OpAmp Adaptation for Zoom Attention to Golden Contexts
    • Divergent Thoughts toward One Goal: LLM-based Multi-Agent Collaboration System for Electronic Design Automation
    • Parameter-Efficient Sparsity Crafting from Dense to Mixture-of-Experts for Instruction Tuning on General Tasks
    • ChatEDA: A Large Language Model Powered Autonomous Agent for EDA
    • p-Laplacian Adaptation for Generative Pre-trained Vision-Language Models
  • Experience

Scaling Native Multimodal Pre-Training From Scratch

Jul 25, 2026·
Haoyuan Wu
Haoyuan Wu
,
Aoqi Wu
,
Hai Wang
,
Jiajia Wu
,
Jinxiang Ou
,
Bei Yu
· 0 min read
Paper
Type
Conference paper
Publication
arXiv:2607.22043 (2026)
Last updated on Jul 25, 2026
Multimodal Models
Haoyuan Wu
Authors
Haoyuan Wu
Ph.D. Student

LLaDA2.0-Uni: Unifying Multimodal Understanding and Generation with Diffusion Large Language Model Apr 22, 2026 →