vix.ing · top · new · best · stats · spec

Imagen 3

2024/08/13 by Imagen-Team-Google, :, Jason Baldridge +526 · 1 voice · 3 citations
Computer Science · Medicine · #Computer Vision and Pattern Recognition (cs.CV) #FOS: Computer and information sciences #Radiomics and Machine Learning in Medical Imaging #cs.CV

paper · pdf · doi:10.48550/arxiv.2408.07009

openalex publication_date 2024/08/13 · arxiv published 2024/08/13 · openalex created_date 2024/09/11 · arxiv updated 2024/12/21 · openalex updated_date 2026/07/28

Abstract

We introduce Imagen 3, a latent diffusion model that generates high quality images from text prompts. We describe our quality and responsibility evaluations. Imagen 3 is preferred over other state-of-the-art (SOTA) models at the time of evaluation. In addition, we discuss issues around safety and representation, as well as methods we used to minimize the potential harm of our models.

Cited by

Discussions

Related