Naoki Murata
- Consistency Trajectory Models: Learning Probability Flow ODE Trajectory of Diffusion
2023/10/01 by Dongjun Kim, Kim, Dongjun, Chieh-Hsin Lai +15 · 137 citations
Computer Science · #Artificial Intelligence (cs.AI) #Computer Vision and Pattern Recognition (cs.CV) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Machine Learning (stat.ML) #Machine Learning and Algorithms #Machine Learning and Data Classification
- Manifold Preserving Guided Diffusion
2023/11/28 by Yutong He, Naoki Murata, He, Yutong +19 · 44 citations
Computer Science · Physics and Astronomy · #Artificial Intelligence (cs.AI) #Computer Vision and Pattern Recognition (cs.CV) #Domain Adaptation and Few-Shot Learning #FOS: Computer and information sciences #Generative Adversarial Networks and Image Synthesis #Machine Learning (cs.LG) #Model Reduction and Neural Networks
- SQ-VAE: Variational Bayes on Discrete Representation with Self-annealed Stochastic Quantization
2022/05/16 by Yuhta Takida, Takashi Shibuya, Takida, Yuhta +17 · 20 citations
Computer Science · #AI in cancer detection #Computer Vision and Pattern Recognition (cs.CV) #Domain Adaptation and Few-Shot Learning #FOS: Computer and information sciences #Machine Learning (cs.LG) #Speech and Audio Processing
- GenWarp: Single Image to Novel Views with Semantic-Preserving Generative Warping
2024/05/27 by Junyoung Seo, Seo, Junyoung, Kazumi Fukuda +15 · 20 citations
Computer Science · #Image Retrieval and Classification Techniques #Generative Adversarial Networks and Image Synthesis #Image Processing and 3D Reconstruction
- GibbsDDRM: A Partially Collapsed Gibbs Sampler for Solving Blind Inverse Problems with Denoising Diffusion Restoration
2023/01/30 by Naoki Murata, Koichi Saito, Murata, Naoki +11 · 11 citations
Computer Science · #Artificial Intelligence (cs.AI) #Audio and Speech Processing (eess.AS) #Computer Vision and Pattern Recognition (cs.CV) #FOS: Computer and information sciences #FOS: Electrical engineering #Generative Adversarial Networks and Image Synthesis #Image and Signal Denoising Methods #Machine Learning (cs.LG) #Music and Audio Processing #Sound (cs.SD) #electronic engineering #information engineering
- MusicMagus: Zero-Shot Text-to-Music Editing via Diffusion Models
2024/02/09 by Yixiao Zhang, Yukara Ikemiya, Zhang, Yixiao +13 · 13 citations
Computer Science · #Music and Audio Processing #Music Technology and Sound Studies #Speech Recognition and Synthesis
- HQ-VAE: Hierarchical Discrete Representation Learning with Variational Bayes
2023/12/31 by Yuhta Takida, Takida, Yuhta, Yukara Ikemiya +19 · 10 citations
Computer Science · Biochemistry, Genetics and Molecular Biology · #Image and Signal Denoising Methods #AI in cancer detection #Cancer-related molecular mechanisms research
- PaGoDA: Progressive Growing of a One-Step Generator from a Low-Resolution Diffusion Teacher
2024/05/23 by Dongjun Kim, Chieh-Hsin Lai, Kim, Dongjun +13 · 9 citations
Computer Science · #Advanced Data Compression Techniques #Artificial Intelligence (cs.AI) #Computer Vision and Pattern Recognition (cs.CV) #Educational Technology and Assessment #FOS: Computer and information sciences #Machine Learning (cs.LG) #Machine Learning (stat.ML)
- FP-Diffusion: Improving Score-based Diffusion Models by Enforcing the Underlying Score Fokker-Planck Equation
2022/10/09 by Chieh-Hsin Lai, Lai, Chieh-Hsin, Yuhta Takida +9 · 5 citations
Computer Science · Mathematics · Physics and Astronomy · #Artificial Intelligence (cs.AI) #FOS: Computer and information sciences #Gaussian Processes and Bayesian Inference #Machine Learning (cs.LG) #Mathematical Biology Tumor Growth #Model Reduction and Neural Networks
- Unsupervised vocal dereverberation with diffusion-based generative models
2022/11/08 by Koichi Saito, Saito, Koichi, Naoki Murata +11 · 4 citations
Computer Science · Engineering · #Speech and Audio Processing #Music and Audio Processing #Acoustic Wave Phenomena Research
- SAN: Inducing Metrizability of GAN with Discriminative Normalized Linear Layer
2023/01/30 by Yuhta Takida, Takida, Yuhta, Masaaki Imaizumi +10 · 5 citations
Computer Science · #AI in cancer detection #FOS: Computer and information sciences #Generative Adversarial Networks and Image Synthesis #Human Pose and Action Recognition #Machine Learning (cs.LG)
- Automated Black-box Prompt Engineering for Personalized Text-to-Image Generation
2024/03/28 by Yutong He, Alexander Robey, He, Yutong +17 · 4 citations
Computer Science · Social Sciences · #Artificial Intelligence (cs.AI) #Computation and Language (cs.CL) #Computer Vision and Pattern Recognition (cs.CV) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Mathematics, Computing, and Information Processing #Multimedia Communication and Technology #Video Analysis and Summarization
- Improving Vector-Quantized Image Modeling with Latent Consistency-Matching Diffusion
2024/10/18 by Nguyễn Hoàng Bắc, Nguyen, Bac, Yuhta Takida +10 · 4 citations
Mathematics · #FOS: Computer and information sciences #Machine Learning (cs.LG) #Statistical Methods and Inference
- SteerMusic: Enhanced Musical Consistency for Zero-shot Text-guided and Personalized Music Editing
2025/04/15 by Xinlei Niu, Niu, Xinlei, Kin Wai Cheuk +19 · 5 citations
Computer Science · #Artificial Intelligence in Games #Audio and Speech Processing (eess.AS) #FOS: Computer and information sciences #FOS: Electrical engineering #Multimedia (cs.MM) #Music Technology and Sound Studies #Music and Audio Processing #Sound (cs.SD) #electronic engineering #information engineering
- Forging and Removing Latent-Noise Diffusion Watermarks Using a Single Image
2025/04/27 by Anubhav Jain, Yuya Kobayashi, Jain, Anubhav +16 · 1 voice · 5 citations
Computer Science · #Advanced Steganography and Watermarking Techniques #Adversarial Robustness in Machine Learning #Computer Vision and Pattern Recognition (cs.CV) #FOS: Computer and information sciences #Generative Adversarial Networks and Image Synthesis #cs.CV
- MoLA: Motion Generation and Editing with Latent Diffusion Enhanced by Adversarial Training
2024/06/04 by Kengo Uchida, Takashi Shibuya, Uchida, Kengo +10 · 2 citations
Computer Science · Engineering · #Computer Vision and Pattern Recognition (cs.CV) #FOS: Computer and information sciences #Generative Adversarial Networks and Image Synthesis #Human Motion and Animation #Human Pose and Action Recognition
- Weighted Point Set Embedding for Multimodal Contrastive Learning Toward Optimal Similarity Metric
2024/04/30 by Toshimitsu Uesaka, Uesaka, Toshimitsu, Taiji Suzuki +9 · 2 citations
Arts and Humanities · Psychology · #EFL/ESL Teaching and Learning #FOS: Computer and information sciences #Innovative Teaching and Learning Methods #Machine Learning (cs.LG)
- G2D2: Gradient-Guided Discrete Diffusion for Inverse Problem Solving
2024/10/09 by Naoki Murata, Chieh-Hsin Lai, Murata, Naoki +11 · 2 citations
Computer Science · Engineering · Mathematics · #Artificial Intelligence (cs.AI) #Computer Vision and Pattern Recognition (cs.CV) #FOS: Computer and information sciences #Image and Signal Denoising Methods #Machine Learning (cs.LG) #Numerical methods in inverse problems #Photoacoustic and Ultrasonic Imaging
- On the Language Encoder of Contrastive Cross-modal Models
2023/10/20 by Mengjie Zhao, Junya Ono, Zhao, Mengjie +17 · 1 citation
Computer Science · Arts and Humanities · #Multimodal Machine Learning Applications #Domain Adaptation and Few-Shot Learning #Subtitles and Audiovisual Media