work FLARE Diffusion for hybrid language models Chimera Designing and Chinchilla-scaling hybrid visual diffusion transformers ADOPD Ongoing work on understanding and generating multimodal documents Context Learning Agentic, fine-grained evaluation of physical and semantic consistency in generated videos fun