vix.ing · top · new · best · stats · spec

Liu, Dongyang

  1. Lumina-mGPT: Illuminate Flexible Photorealistic Text-to-Image Generation with Multimodal Generative Pretraining
    2024/08/05 by Dongyang Liu, Shitian Zhao, Liu, Dongyang +12 · 34 citations
    Engineering · Computer Science · Health Professions · #Human Motion and Animation #Multimodal Machine Learning Applications #Digital Storytelling and Education
  2. Lumina-Next: Making Lumina-T2X Stronger and Faster with Next-DiT
    2024/06/05 by Le Zhuo, Ruoyi Du, Zhuo, Le +41 · 31 citations
    Computer Science · Physics and Astronomy · #Generative Adversarial Networks and Image Synthesis #Music Technology and Sound Studies #Model Reduction and Neural Networks
  3. Z-Image: An Efficient Image Generation Foundation Model with Single-Stream Diffusion Transformer
    2025/11/27 by Z-Image Team, Huanqia Cai, Image Team +46 · 2 voices · 17 citations
    Computer Science · #Advanced Neural Network Applications #Generative Adversarial Networks and Image Synthesis #Image Enhancement Techniques #cs.CV
  4. Lumina-Image 2.0: A Unified and Efficient Image Generative Framework
    2025/03/27 by Qin, Qi, Zhuo, Le, Xin, Yi +20 · 32 citations
    #Computer Vision and Pattern Recognition (cs.CV) #FOS: Computer and information sciences
  5. Lumina-T2X: Transforming Text into Any Modality, Resolution, and Duration via Flow-based Large Diffusion Transformers
    2024/05/09 by Gao, Peng, Zhuo, Le, Liu, Dongyang +17 · 14 citations
    #Computer Vision and Pattern Recognition (cs.CV) #FOS: Computer and information sciences
  6. VEnhancer: Generative Space-Time Enhancement for Video Generation
    2024/07/10 by Jingwen He, Tianfan Xue, He, Jingwen +15 · 14 citations
    Computer Science · #Advanced Vision and Imaging #Computer Graphics and Visualization Techniques #Generative Adversarial Networks and Image Synthesis
  7. SPHINX-X: Scaling Data and Parameters for a Family of Multi-modal Large Language Models
    2024/02/08 by Renrui Zhang, Liu, Dongyang, Zhang, Renrui +34 · 11 citations
    Computer Science · #Artificial Intelligence (cs.AI) #Computation and Language (cs.CL) #Computer Vision and Pattern Recognition (cs.CV) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Natural Language Processing Techniques #Topic Modeling
  8. I-Max: Maximize the Resolution Potential of Pre-trained Rectified Flow Transformers with Projected Flow
    2024/10/10 by Du, Ruoyi, Liu, Dongyang, Zhuo, Le +4 · 10 citations
    #Computer Vision and Pattern Recognition (cs.CV) #FOS: Computer and information sciences
  9. Lumina-mGPT 2.0: Stand-Alone AutoRegressive Image Modeling
    2025/07/23 by Xin Yi, Juncheng Yan, Xin, Yi +37 · 21 citations
    Computer Science · Engineering · Neuroscience · #Brain Tumor Detection and Classification #Computer Vision and Pattern Recognition (cs.CV) #FOS: Computer and information sciences #Medical Image Segmentation Techniques #Medical Imaging and Analysis
  10. Give Me Something to Eat: Referring Expression Comprehension with Commonsense Knowledge
    2020/06/02 by Wang, Peng, Liu, Dongyang, Li, Hui +1 · 2 citations
    #Computer Vision and Pattern Recognition (cs.CV) #FOS: Computer and information sciences
  11. Function-Consistent Feature Distillation
    2023/04/24 by Liu, Dongyang, Kan, Meina, Shan, Shiguang +1 · 3 citations
    #Computer Vision and Pattern Recognition (cs.CV) #FOS: Computer and information sciences
  12. Lumina-Video: Efficient and Flexible Video Generation with Multi-scale Next-DiT
    2025/02/10 by Dongyang Liu, Shicheng Li, Liu, Dongyang +35 · 4 citations
    Computer Science · #Advanced Vision and Imaging #Computer Vision and Pattern Recognition (cs.CV) #FOS: Computer and information sciences #Image and Video Quality Assessment #Video Coding and Compression Technologies
  13. Decoupled DMD: CFG Augmentation as the Spear, Distribution Matching as the Shield
    2025/11/27 by Dongyang Liu, Peng Gao, Liu, Dongyang +18 · 7 citations
    Computer Science · #Generative Adversarial Networks and Image Synthesis #Advanced Neural Network Applications #Advanced Image and Video Retrieval Techniques
  14. OmniCaptioner: One Captioner to Rule Them All
    2025/04/09 by Yiting Lu, Jiakang Yuan, Lu, Yiting +36 · 5 citations
    Computer Science · #Multimodal Machine Learning Applications #Generative Adversarial Networks and Image Synthesis #Data Visualization and Analytics
  15. Distribution Matching Distillation Meets Reinforcement Learning
    2025/11/17 by Dengyang Jiang, Dongyang Liu, Jiang, Dengyang +21 · 5 citations
    Computer Science · #Domain Adaptation and Few-Shot Learning #Generative Adversarial Networks and Image Synthesis #Image Enhancement Techniques
  16. LeX-Art: Rethinking Text Generation via Scalable High-Quality Data Synthesis
    2025/03/27 by Shitian Zhao, Qilong Wu, Zhao, Shitian +23 · 3 citations
    Computer Science · #Computer Vision and Pattern Recognition (cs.CV) #FOS: Computer and information sciences #Image Processing and 3D Reconstruction #Mathematics, Computing, and Information Processing #Natural Language Processing Techniques