work FLARE Diffusion for hybrid language models Chimera Designing and Chinchilla-scaling hybrid visual diffusion transformers ADOPD (2024, 2026) Ongoing work on understanding and generating multimodal documents Context Learning Agentic, fine-grained evaluation of physical and semantic consistency in generated videos fun