Schroff, Florian
- Rethinking Atrous Convolution for Semantic Image Segmentation
2017/06/17 by Liang-Chieh Chen, Chen, Liang-Chieh, George Papandreou +5 · 330 citations
Computer Science · #Advanced Image and Video Retrieval Techniques #Advanced Neural Network Applications #Computer Vision and Pattern Recognition (cs.CV) #FOS: Computer and information sciences #Image Retrieval and Classification Techniques
- Encoder-Decoder with Atrous Separable Convolution for Semantic Image Segmentation
2018/02/07 by Chen, Liang-Chieh, Zhu, Yukun, Papandreou, George +2 · 232 citations
#Computer Vision and Pattern Recognition (cs.CV) #FOS: Computer and information sciences
- VideoPrism: A Foundational Visual Encoder for Video Understanding
2024/02/20 by Long Zhao, Zhao, Long, Nitesh B. Gundavarapu +35 · 2 voices · 22 citations
Computer Science · #Artificial Intelligence (cs.AI) #Computer Vision and Pattern Recognition (cs.CV) #FOS: Computer and information sciences #cs.AI #cs.CV
- Searching for Efficient Multi-Scale Architectures for Dense Image Prediction
2018/09/11 by Liang-Chieh Chen, Chen, Liang-Chieh, Maxwell D. Collins +13 · 50 citations
Computer Science · #Domain Adaptation and Few-Shot Learning #Advanced Image and Video Retrieval Techniques #Advanced Neural Network Applications
- Modeling Uncertainty with Hedged Instance Embedding
2018/09/30 by Seong Joon Oh, Oh, Seong Joon, Kevin Murphy +9 · 6 citations
Computer Science · #AI in cancer detection #Computer Vision and Pattern Recognition (cs.CV) #Domain Adaptation and Few-Shot Learning #FOS: Computer and information sciences #Generative Adversarial Networks and Image Synthesis #Machine Learning (cs.LG) #Machine Learning (stat.ML)
- Imagen 3
2024/08/13 by Imagen-Team-Google, Jason Baldridge, : +526 · 1 voice · 3 citations
Computer Science · Medicine · #Computer Vision and Pattern Recognition (cs.CV) #FOS: Computer and information sciences #Radiomics and Machine Learning in Medical Imaging #cs.CV
- Auto-DeepLab: Hierarchical Neural Architecture Search for Semantic Image Segmentation
2019/01/10 by Liu, Chenxi, Chen, Liang-Chieh, Schroff, Florian +4 · 5 citations
#Computer Vision and Pattern Recognition (cs.CV) #FOS: Computer and information sciences #Machine Learning (cs.LG)
- Distilling Vision-Language Models on Millions of Videos
2024/01/11 by Yue Zhao, Long Zhao, Zhao, Yue +22 · 1 voice · 2 citations
Computer Science · #Multimodal Machine Learning Applications #Natural Language Processing Techniques #Topic Modeling
- View-Invariant Probabilistic Embedding for Human Pose
2019/12/02 by Jennifer J. Sun, Jiaping Zhao, Sun, Jennifer J. +9 · 3 citations
Computer Science · Engineering · #Computer Vision and Pattern Recognition (cs.CV) #FOS: Computer and information sciences #Gait Recognition and Analysis #Human Pose and Action Recognition #Video Surveillance and Tracking Methods
- Unified Visual Relationship Detection with Vision and Language Models
2023/03/16 by Zhao, Long, Yuan, Liangzhe, Gong, Boqing +5 · 3 citations
#Computer Vision and Pattern Recognition (cs.CV) #FOS: Computer and information sciences
- Learning View-Disentangled Human Pose Representation by Contrastive Cross-View Mutual Information Maximization
2020/12/02 by Zhao, Long, Wang, Yuxiao, Zhao, Jiaping +7 · 1 citation
#Computer Vision and Pattern Recognition (cs.CV) #FOS: Computer and information sciences
- Contextualized Spatio-Temporal Contrastive Learning with Self-Supervision
2021/12/09 by Yuan, Liangzhe, Qian, Rui, Cui, Yin +5 · 1 citation
#Computer Vision and Pattern Recognition (cs.CV) #FOS: Computer and information sciences
- Structured Video-Language Modeling with Temporal Grouping and Spatial Grounding
2023/03/28 by Xiong, Yuanhao, Zhao, Long, Gong, Boqing +5 · 1 citation
#Computer Vision and Pattern Recognition (cs.CV) #FOS: Computer and information sciences
- VideoGLUE: Video General Understanding Evaluation of Foundation Models
2023/07/06 by Liangzhe Yuan, Yuan, Liangzhe, Nitesh B. Gundavarapu +31 · 1 citation
Computer Science · #Human Pose and Action Recognition #Multimodal Machine Learning Applications #Domain Adaptation and Few-Shot Learning