2017/05/22 by Tony Beltramelli, Beltramelli, Tony · 4 voices · 28 citations
Computer Science · #Advanced Image and Video Retrieval Techniques #Multimodal Machine Learning Applications #Video Analysis and Summarization #cs.AI #cs.CL #cs.CV #cs.LG #cs.NE
paper · pdf · doi:10.48550/arxiv.1705.07962
openalex publication_date 2017/05/22 · openalex created_date 2025/10/10 · openalex updated_date 2026/07/28
Transforming a graphical user interface screenshot created by a designer into computer code is a typical task conducted by a developer in order to build customized software, websites, and mobile applications. In this paper, we show that deep learning methods can be leveraged to train a model end-to-end to automatically generate code from a single input image with over 77% of accuracy for three different platforms (i.e. iOS, Android and web-based technologies).