YFCC100M
2016/01/25 by Bart Thomee, Bart Thomée, David A. Shamma +7 · 107 citations
Biochemistry, Genetics and Molecular Biology · Medicine · #Cancer Genomics and Diagnostics #Radiomics and Machine Learning in Medical Imaging
paper · pdf · doi:10.1145/2812802
Abstract
This publicly available curated dataset of almost 100 million photos and videos is free and legal for all.
Citations
Cited by
- Digital collections explorer: An open-source, multimodal viewer for searching digital collections
- ILIAS: Instance-Level Image retrieval At Scale
- Procrustean Training for Imbalanced Deep Learning
- INTERN: A New Learning Paradigm Towards General Vision
- MobileNets: Efficient Convolutional Neural Networks for Mobile Vision Applications
- Demystifying Contrastive Self-Supervised Learning: Invariances, Augmentations and Dataset Biases
- Stochastic Neighbor Embedding of Multimodal Relational Data for Image-Text Simultaneous Visualization
- Multimodal datasets: misogyny, pornography, and malignant stereotypes
- Towards Generative Location Awareness for Disaster Response: A Probabilistic Cross-view Geolocalization Approach
- MVImgNet2.0: A Larger-scale Dataset of Multi-view Images
- Patch-wise Retrieval: A Bag of Practical Techniques for Instance-level Matching
- AdaLAM: Revisiting Handcrafted Outlier Detection
- MUSIQ: Multi-scale Image Quality Transformer
- ContextDesc: Local Descriptor Augmentation with Cross-Modality Context
- Momentum Contrast for Unsupervised Visual Representation Learning
- Learning Transferable Visual Models From Natural Language Supervision
- Image recognition from raw labels collected without annotators
- Image Matching from Handcrafted to Deep Features: A Survey
- Progressive Correspondence Pruning by Consensus Learning
- Learning Generalized Spatial-Temporal Deep Feature Representation for No-Reference Video Quality Assessment
- Certified Data Removal from Machine Learning Models
- Learning from Noisy Labels with Distillation
- Training and Evaluating Multimodal Word Embeddings with Large-scale Web Annotated Images
- Tackling the Problem of Limited Data and Annotations in Semantic\n Segmentation
- Learning Two-View Correspondences and Geometry Using Order-Aware Network
- Changing Fashion Cultures
- GeoWINE: Geolocation based Wiki, Image,News and Event Retrieval
- CosFace: Large Margin Cosine Loss for Deep Face Recognition
- Dual-Glance Model for Deciphering Social Relationships
- LeCoT: revisiting network architecture for two-view correspondence pruning
- Image Matching Across Wide Baselines: From Paper to Practice
- GeoToken: Hierarchical Geolocalization of Images via Next Token Prediction
- NExT-QA:Next Phase of Question-Answering to Explaining Temporal Actions
- Learning Generalized Spatial-Temporal Deep Feature Representation for No-Reference Video Quality Assessment
- DisCoVQA: Temporal Distortion-Content Transformers for Video Quality Assessment
- LMM-VQA: Advancing Video Quality Assessment With Large Multimodal Models
- Online Continual Learning with Natural Distribution Shifts: An Empirical Study with Visual Data
- Graph convolutional networks for learning with few clean and many noisy labels
- Revisiting IM2GPS in the Deep Learning Era
- Geo-Aware Networks for Fine-Grained Recognition
- Real-Time Adaptive Image Compression
- From Photo Streams to Evolving Situations
- Subjective and Objective Audio-Visual Quality Assessment for User Generated Content
- Data-Efficient Image Recognition with Contrastive Predictive Coding
- Study of Spatio-Temporal Modeling in Video Quality Assessment
- WebFace260M: A Benchmark Unveiling the Power of Million-Scale Deep Face Recognition
- Richard Hooker on the eucharist : A commentary on the Laws V.67
- Self-training with Noisy Student improves ImageNet classification
- PixelTransformer: Sample Conditioned Signal Generation
- Zero-Shot Text-to-Image Generation
- Improving On-Screen Sound Separation for Open-Domain Videos with\n Audio-Visual Self-Attention
- Multilevel Language and Vision Integration for Text-to-Clip Retrieval
- Image Quality Assessment Using Contrastive Learning
- Generative AI in Depth: A Survey of Recent Advances, Model Variants, and Real-World Applications
- Learning Imbalanced Datasets with Label-Distribution-Aware Margin Loss
- PPR-FCN: Weakly Supervised Visual Relation Detection via Parallel Pairwise R-FCN
- Momentum Contrast for Unsupervised Visual Representation Learning
- The Hateful Memes Challenge: Detecting Hate Speech in Multimodal Memes
- Tag Prediction at Flickr: a View from the Darkroom
- Visual Relationship Detection using Scene Graphs: A Survey
- Compact deep neural network models of the visual cortex
- The Photographic Pipeline of Machine Vision; or, Machine Vision's Latent Photographic Theory
- YOLO9000: Better, Faster, Stronger
- Identifying and Compensating for Feature Deviation in Imbalanced Deep Learning
- Trade-offs in Cross-Domain Generalization of Foundation Model Fine-Tuned for Biometric Applications
- Vision-Language Models for Vision Tasks: A Survey
- A Self-Explainable Stylish Image Captioning Framework via Multi-References
- XCiT: Cross-Covariance Image Transformers
- MultiGrain: a unified image embedding for classes and instances
- Emerging Properties in Self-Supervised Vision Transformers
- Effectively obtaining acoustic, visual and textual data from videos
- Unsupervised Instance Segmentation with Superpixels
- Supervision Exists Everywhere: A Data Efficient Contrastive Language-Image Pre-training Paradigm
- Radioactive data: tracing through training
- SODA10M: A Large-Scale 2D Self/Semi-Supervised Object Detection Dataset for Autonomous Driving
- How to Train Your MAML to Excel in Few-Shot Classification
- CLOOB: Modern Hopfield Networks with InfoLOOB Outperform CLIP
- FILIP: Fine-grained Interactive Language-Image Pre-Training
- Channel-Level Variable Quantization Network for Deep Image Compression
- S2M3: Split-and-Share Multi-Modal Models for Distributed Multi-Task Inference on the Edge
- Probabilistic Video Generation using Holistic Attribute Control
- Déjà Vu: an empirical evaluation of the memorization properties of ConvNets
- Object Recognition Datasets and Challenges: A Review
- Predicting floods with Flickr tags. [europepmc]
- Global multi-layer network of human mobility. [europepmc]
- The Socio-Moral Image Database (SMID): A novel stimulus set for the study of social, moral and affective processes. [europepmc]
- Using social media to quantify spatial and temporal dynamics of nature-based recreational activities. [europepmc]
- QUADRIVEN: A Framework for Qualitative Taxi Demand Prediction Based on Time-Variant Online Social Network Data Analysis. [europepmc]
- No Reference, Opinion Unaware Image Quality Assessment by Anomaly Detection. [europepmc]
- FASDQ: Fault-Tolerant Adaptive Scheduling with Dynamic QoS-Awareness in Edge Containers for Delay-Sensitive Tasks. [europepmc]
- Toward Learning Trustworthily from Data Combining Privacy, Fairness, and Explainability: An Application to Face Recognition. [europepmc]
- No-Reference Image Quality Assessment with Global Statistical Features. [europepmc]
- No-Reference Quality Assessment of In-Capture Distorted Videos. [europepmc]
- An Efficient Method for No-Reference Video Quality Assessment. [europepmc]
- No-Reference Video Quality Assessment Using Multi-Pooled, Saliency Weighted Deep Features and Decision Fusion. [europepmc]
- AdaSG: A Lightweight Feature Point Matching Method Using Adaptive Descriptor with GNN for VSLAM. [europepmc]
- CLIP knows image aesthetics. [europepmc]
- No-Reference Video Quality Assessment Using the Temporal Statistics of Global and Local Image Features. [europepmc]
- RANSAC for Robotic Applications: A Survey. [europepmc]
- Conv-Former: A Novel Network Combining Convolution and Self-Attention for Image Quality Assessment. [europepmc]
- Generalization of vision pre-trained models for histopathology. [europepmc]
- Using HVS Dual-Pathway and Contrast Sensitivity to Blindly Assess Image Quality. [europepmc]
- Analysis and interpretation of joint source separation and sound event detection in domestic environments. [europepmc]
- No-Reference Image Quality Assessment with Multi-Scale Orderless Pooling of Deep Features. [europepmc]
- Multimodal data integration for oncology in the era of deep neural networks: a review. [europepmc]
- Significantly improving zero-shot X-ray pathology classification via fine-tuning pre-trained image-text encoders. [europepmc]
- Leveraging two-dimensional pre-trained vision transformers for three-dimensional model generation via masked autoencoders. [europepmc]
- Fair human-centric image dataset for ethical AI benchmarking. [europepmc]
Related