ImageNet classification with deep convolutional neural networks
2017/05/24 by Alex Krizhevsky, Ilya Sutskever, Geoffrey E. Hinton · 75,716 citations
Computer Science · #Advanced Image Processing Techniques #Advanced Neural Network Applications #Artificial intelligence #Artificial neural network #Computer science #Convolution (computer science) #Convolutional neural network #Deep neural networks #Domain Adaptation and Few-Shot Learning #Dropout (neural networks) #Machine learning #Normalization (sociology) #Pattern recognition (psychology) #Pooling #Regularization (linguistics) #Softmax function #Word error rate
paper · pdf · doi:10.1145/3065386
published in Communications of the ACM 60(6), 84-90 (Association for Computing Machinery)
openalex publication_date 2017/05/24 · openalex created_date 2025/10/10 · openalex updated_date 2026/08/01
Abstract
We trained a large, deep convolutional neural network to classify the 1.2 million high-resolution images in the ImageNet LSVRC-2010 contest into the 1000 different classes. On the test data, we achieved top-1 and top-5 error rates of 37.5% and 17.0%, respectively, which is considerably better than the previous state-of-the-art. The neural network, which has 60 million parameters and 650,000 neurons, consists of five convolutional layers, some of which are followed by max-pooling layers, and three fully connected layers with a final 1000-way softmax. To make training faster, we used non-saturating neurons and a very efficient GPU implementation of the convolution operation. To reduce overfitting in the fully connected layers we employed a recently developed regularization method called "dropout" that proved to be very effective. We also entered a variant of this model in the ILSVRC-2012 competition and achieved a winning top-5 test error rate of 15.3%, compared to 26.2% achieved by the second-best entry.
Cited by
- Generative Adversarial Networks: A Survey Towards Private and Secure Applications
- Data-Driven Aerospace Engineering: Reframing the Industry with Machine Learning
- YFCC100M
- Convolutional neural network for seismic impedance inversion
- Elastic prestack seismic inversion through discrete cosine transform reparameterization and convolutional neural networks
- Deep Learning Methods for Reynolds-Averaged Navier-Stokes Simulations of Airfoil Flows
- Deep Learning with Coherent Nanophotonic Circuits
- Recent Advances and Applications of Machine Learning in Experimental Solid Mechanics: A Review
- Quantum-chemical insights from deep tensor neural networks
- Using Deep Learning and Google Street View to Estimate the Demographic Makeup of the US
- Artificial Intelligence in manufacturing: State of the art, perspectives, and future directions
- The Principles of Deep Learning Theory
- A Comprehensive Survey on Transfer Learning
- Edge Intelligence: Paving the Last Mile of Artificial Intelligence With Edge Computing
- Tumour-infiltrating lymphocytes: from prognosis to treatment selection
- Applications and Techniques for Fast Machine Learning in Science
- Multiclass magnetic resonance imaging brain tumor classification using artificial intelligence paradigm
- Prediction of causative genes in inherited retinal disorder from fundus photography and autofluorescence imaging using deep learning techniques
- Toward Causal Representation Learning
- Deep Reinforcement Learning Based Resource Allocation for V2V Communications
- Real-time differentiation of adenomatous and hyperplastic diminutive colorectal polyps during analysis of unaltered videos of standard colonoscopy using a deep learning model
- Artificial intelligence in fetal brain imaging: Advancements, challenges, and multimodal approaches for biometric and structural analysis
- Deep convolutional neural network for the automated detection and diagnosis of seizure using EEG signals
- Obstructive sleep apnea detection from single-lead electrocardiogram signals using one-dimensional squeeze-and-excitation residual group network
- A scoping review of transfer learning research on medical image analysis using ImageNet
- Advancements in automated nuclei segmentation for histopathology using you only look once-driven approaches: A systematic review
- A novel wavelet sequence based on deep bidirectional LSTM network model for ECG signal classification
- Transparency of deep neural networks for medical image analysis: A review of interpretability methods
- Deep learning for denoising
- Regularized elastic full-waveform inversion using deep learning
- Deep learning for forest inventory and planning: a critical review on the remote sensing approaches so far and prospects for further applications
- Deep Neural Network Approximation Theory
- Deep learning and its application in geochemical mapping
- PYPM-GGD: Pitman-Yor Process Mixture with Generalized Gaussian Density using ADAM
- DSIS - a database system with interrelational semantics
- Physics and Equality Constrained Artificial Neural Networks: Application to Forward and Inverse Problems with Multi-fidelity Data Fusion
- DeepWalk
- Dominant-Current Deep Learning Scheme for Electrical Impedance Tomography
- Efficient Mitchell’s Approximate Log Multipliers for Convolutional Neural Networks
- MTHAEL: Cross-Architecture IoT Malware Detection Based on Neural Network Advanced Ensemble Learning
- Efficient Processing of Deep Neural Networks: A Tutorial and Survey
- Classification of the Clinical Images for Benign and Malignant Cutaneous Tumors Using a Deep Learning Algorithm
- Deep Learning With Edge Computing: A Review
- Neuro-Inspired Computing With Emerging Nonvolatile Memorys
- FeatureNet: Machining feature recognition based on 3D Convolution Neural Network
- Automated arrhythmia detection with homeomorphically irreducible tree technique using more than 10,000 individual subject ECG records
- Dual non-autonomous deep convolutional neural network for image denoising
- Image-Based Multiresolution Topology Optimization Using Deep Disjunctive Normal Shape Model
- Physics-informed neural networks: A deep learning framework for solving forward and inverse problems involving nonlinear partial differential equations
- Optimal control of PDEs using physics-informed neural networks
- Characterising the neural time-courses of food attribute representations
- Deep Learning for Economists
- LiMuon: Light and Fast Muon Optimizer for Large Models
- AdamNX: An Adam improvement algorithm based on a novel exponential decay mechanism for the second-order moment estimate
- Reclaiming AI as a Theoretical Tool for Cognitive Science
- Hospital Length of Stay Prediction Methods
- Hypergraph convolution and hypergraph attention
- Rasabodha: Understanding Indian classical dance by recognizing emotions using deep learning
- U2-Net: Going Deeper with Nested U-Structure for Salient Object Detection
- Multi-crop Convolutional Neural Networks for lung nodule malignancy suspiciousness classification
- Illumination-aware faster R-CNN for robust multispectral pedestrian detection
- Haar wavelet downsampling: A simple but effective downsampling module for semantic segmentation
- Watch, attend and parse: An end-to-end neural network based approach to handwritten mathematical expression recognition
- Introduction to deep learning methods for multi‐species predictions
- Deep photovoltaic nowcasting
- Wear particle classification considering particle overlapping
- Abstraction-Based Proof Production in Formal Verification of Neural Networks
- Automatic morphologic classification of Martian craters using imbalanced datasets of Tianwen-1’s MoRIC images with deep neural networks
- Anomalous-diffusion synthesis of non-Gaussian reservoir anomalies for time-lapse seismic inversion
- Exact Neural-Network Representations of the Motzkin States
- Conservative physics-informed neural networks on discrete domains for conservation laws: Applications to forward and inverse problems
- Class-Balanced Softmax: A Bayes Theory-Based Method for Long-Tailed Recognition
- A novel deep learning-based modelling strategy from image of particles to mechanical properties for granular materials with CNN and BiLSTM
- Wireless Networks Design in the Era of Deep Learning: Model-Based, AI-Based, or Both?
- Deep learning for the partially linear Cox model
- Real-Time Detection of Charge Jumps in Superconducting Qubits with a Convolutional Neural Network
- 4DGS360: 360° Gaussian Reconstruction of Dynamic Objects from a Single Video
- Pointing-Based Object Recognition
- Prediction model of temperature field in dual-mode combustors based on wall pressure
- Characterization of heat release rate by OH* and CH* chemiluminescence
- Evaluating Uncertainty and Quality of Visual Language Action-enabled Robots
- SEED: Towards More Accurate Semantic Evaluation for Visual Brain Decoding
- Glyce: Glyph-vectors for Chinese Character Representations
- Scaling Synthetic-Image Pre-Training for Federated Fine-Tuning of Large Vision Models
- MindPilot: Closed-loop Visual Stimulation Optimization for Brain Modulation with EEG-guided Diffusion
- Norm or Direction? Decoding Vision Mambas for High-Resolution Vision
- LieBN: Batch Normalization over Lie Groups
- Edge-Local and Qubit-Efficient Quantum Graph Learning for the NISQ Era
- Benchmarking NACTI Species Recognition in Long-Tailed Regimes
- The Many Senses of Visual Similarity: A Text-Prompted Image Perceptual Metric
- EVOLVE: Efficient Learned Volume Compression with Variable-Rate Encoding on a Cross-Domain Database
- UMCP: A Unified Multi-Task Collaborative Perception Network for Luggage Trolley Pose Estimation
- Provably Lossless Acceleration of DNN Mutation Testing via Memoization
- Direct Clinical Joint Angle Extraction from Parametric Body Model Rotation Matrices
- Artificially intelligent agents in the social and behavioral sciences: A history and outlook
- Feature-Guided Diffusion for Non-Differentiable Inverse Rendering
- VecFontLLM: Anchor-Guided Direct Synthesis of Chinese Vector Fonts
- Dropout and Random Gradient Masking Are Asymptotically Equivalent in Large ResNets
- In-context learning of closed form solution to simple linear regression task using transformer with linear self-attention
- AIMS: An uncertainty-aware AI experimentalist for quantum matter
- Cotton-SF YOLO: Learning Structural and Frequency Cues for Early Cotton Square Detection in Complex Field Environments
- Explicit Over Implicit: Enhancing CNNs Via Complex Structure Tensor Representations for Periocular Recognition
- CNN-Based Surface Temperature Forecasts with Ensemble Numerical Weather Prediction
- Information Theory and Statistical Learning
- Do Machines Fail Like Humans? A Human-Centred Out-of-Distribution Spectrum for Mapping Error Alignment
- AI solutions for evolutionary genomics of nonmodel species
- BehaveAI enables rapid detection and classification of objects and behavior from motion
- The Acceleration of Artificial Intelligence: Rethinking Organization and Work in an Era of Rapid Technological Change
- The Universal Weight Subspace Hypothesis
- MRD: Using Physically Based Differentiable Rendering to Probe Vision Models for 3D Scene Understanding
- Artificial intelligence for risk assessment and outcome prediction in malignant haematology
- Impact of Multi-View Fusion and Biomechanical Modeling on Markerless Motion Tracking
- Living Synthetic Benchmarks: A Neutral and Cumulative Framework for Simulation Studies
- FlyTrap: Physical Distance-Pulling Attack Towards Camera-based Autonomous Target Tracking Systems
- Video models are zero-shot learners and reasoners
- ImageNet-trained CNNs are not biased towards texture: Revisiting feature reliance through controlled suppression
- The Temporal Scaffolding of Sensory Organization
- scPortrait integrates single-cell images into multimodal modeling
- uGMM-NN: Univariate Gaussian Mixture Model Neural Network
- Power Stabilization for AI Training Datacenters
- Neuromorphic Computing: A Theoretical Framework for Time, Space, and Energy Scaling
- Optimizers Qualitatively Alter Solutions And We Should Leverage This
- Dynamic Chunking for End-to-End Hierarchical Sequence Modeling
- Potential role of developmental experience in the emergence of the parvo-magno distinction
- Are Statistical Methods Obsolete in the Era of Deep Learning? A Study of ODE Inverse Problems
- Perception Encoder: The best visual embeddings are not at the output of the network
- Brain-guided convolutional neural networks reveal task-specific representations in scene processing
- Improving Computer Vision Interpretability: Transparent Two-Level Classification for Complex Scenes
- What Is Artificial General Intelligence?
- Deep Learning is Not So Mysterious or Different
- Can machines learn density functionals? Past, present, and future of ML in DFT
- ILIAS: Instance-Level Image retrieval At Scale
- Scaling Laws in Patchification: An Image Is Worth 50,176 Tokens And More
- Evolution and The Knightian Blindspot of Machine Learning
- Language models and Automated Essay Scoring
- Representations of Sound in Deep Learning of Audio Features from Music
- FedRAD: Federated Robust Adaptive Distillation
- Controllable Data Augmentation Through Deep Relighting
- A Convolutional Attention Network for Extreme Summarization of Source Code
- Modeling Spatial and Temporal Cues for Multi-label Facial Action Unit Detection
- Reducing Data Motion to Accelerate the Training of Deep Neural Networks
- Computer Vision and Abnormal Patient Gait Assessment a Comparison of Machine Learning Models
- TorchQuantumDistributed
- Omni-Attribute: Open-vocabulary Attribute Encoder for Visual Concept Personalization
- Profiling based Out-of-core Hybrid Method for Large Neural Networks
- How Deep is the Feature Analysis underlying Rapid Visual Categorization?
- Sampled Softmax with Random Fourier Features
- Spatially Correlated Patterns in Adversarial Images
- Learning to Compose Hypercolumns for Visual Correspondence
- Machine Translation: A Literature Review
- Deep Air Quality Forecasting Using Hybrid Deep Learning Framework
- Procrustean Training for Imbalanced Deep Learning
- NullSpaceNet: Nullspace Convoluional Neural Network with Differentiable Loss Function
- A Convolutional Neural Network for gaze preference detection: A potential tool for diagnostics of autism spectrum disorder in children
- Building Compact and Robust Deep Neural Networks with Toeplitz Matrices
- RotNet: Fast and Scalable Estimation of Stellar Rotation Periods Using Convolutional Neural Networks
- Meta Cross-Modal Hashing on Long-Tailed Data
- Matrix Smoothing: A Regularization for DNN with Transition Matrix under Noisy Labels
- Speeding up convolutional networks pruning with coarse ranking
- Generating Unrestricted 3D Adversarial Point Clouds
- Is Heterophily A Real Nightmare For Graph Neural Networks To Do Node Classification?
- Adaptive Periodic Averaging: A Practical Approach to Reducing Communication in Distributed Learning
- Differentiable Learning-to-Normalize via Switchable Normalization
- NAS-DIP: Learning Deep Image Prior with Neural Architecture Search
- A Simple Saliency Method That Passes the Sanity Checks
- SoK: How Robust is Image Classification Deep Neural Network Watermarking? (Extended Version)
- Deep Multi-View Spatial-Temporal Network for Taxi Demand Prediction
- Neural Abstractive Text Summarization with Sequence-to-Sequence Models
- MUTE: Data-Similarity Driven Multi-hot Target Encoding for Neural Network Design
- Generalized Focal Loss V2: Learning Reliable Localization Quality Estimation for Dense Object Detection
- Multi-Subspace Neural Network for Image Recognition
- Decentralized Deep Reinforcement Learning for Network Level Traffic Signal Control
- A Visual Analytics Framework for Explaining and Diagnosing Transfer Learning Processes
- Colorectal cancer diagnosis from histology images: A comparative study
- Norm-based generalisation bounds for multi-class convolutional neural networks
- Universal Approximation with Quadratic Deep Networks
- Unconstrained Road Marking Recognition with Generative Adversarial Networks
- IBM Deep Learning Service
- Fast GPU Linear Algebra via Compile Time Expression Fusion
- Rethinking Intrinsic Dimension Estimation in Neural Representations
- Evolving Deep Convolutional Neural Networks for Image Classification
- Predictive Analysis of COVID-19 Time-series Data from Johns Hopkins University
- Per-pixel Classification Rebar Exposures in Bridge Eye-inspection
- Deep and Shallow Covariance Feature Quantization for 3D Facial Expression Recognition
- Manifold Criterion Guided Transfer Learning via Intermediate Domain Generation
- Gated Feedback Refinement Network for Coarse-to-Fine Dense Semantic Image Labeling
- TopoResNet: A hybrid deep learning architecture and its application to skin lesion classification
- PCR-ORB: Enhanced ORB-SLAM3 with Point Cloud Refinement Using Deep Learning-Based Dynamic Object Filtering
- End to End Learning for Self-Driving Cars
- Open Problems in Cooperative AI
- Sparse R-CNN: End-to-End Object Detection with Learnable Proposals
- Deep Learning for Needle Detection in a Cannulation Simulator
- The Dynamics of Gradient Descent for Overparametrized Neural Networks
- Progressive Learning of Low-Precision Networks
- Why do linear SVMs trained on HOG features perform so well?
- Scaling Wide Residual Networks for Panoptic Segmentation
- Learning Hybrid Representation by Robust Dictionary Learning in Factorized Compressed Space
- Deep Learning for Spectrum Sensing
- Energy and Memory-Efficient Federated Learning With Ordered Layer Freezing
- Enhancing Convolutional Neural Networks for Face Recognition with Occlusion Maps and Batch Triplet Loss
- Parallel Support Vector Machines in Practice
- A Benchmarking Framework for Interactive 3D Applications in the Cloud
- When science meets geopolitics: global AI research network transformation (2000–2025)
- VideoFlow: A Conditional Flow-Based Model for Stochastic Video Generation
- CNN-based Automatic Detection of Bone Conditions via Diagnostic CT Images for Osteoporosis Screening
- Potentials and challenges of polymer informatics: exploiting machine learning for polymer design
- RMP-SNN: Residual Membrane Potential Neuron for Enabling Deeper High-Accuracy and Low-Latency Spiking Neural Network
- Perceiver: General Perception with Iterative Attention
- Alleviating Bottlenecks for DNN Execution on GPUs via Opportunistic Computing
- Learning a Reinforced Agent for Flexible Exposure Bracketing Selection
- Solving Optical Tomography with Deep Learning
- The Challenge of Multi-Operand Adders in CNNs on FPGAs: How not to solve it!
- ExAD: An Ensemble Approach for Explanation-based Adversarial Detection
- With Great Context Comes Great Prediction Power: Classifying Objects via Geo-Semantic Scene Graphs
- A Simple Method for Commonsense Reasoning
- Alignment of electron optical beam shaping elements using a convolutional neural network
- Partial Graph Reasoning for Neural Network Regularization
- A mountable toilet system for personalized health monitoring via the analysis of excreta
- Immersive exposure to simulated visual hallucinations modulates high-level human cognition
- ChebLieNet: Invariant Spectral Graph NNs Turned Equivariant by Riemannian Geometry on Lie Groups
- Semi-supervised Sequence Learning
- MLP-Mixer: An all-MLP Architecture for Vision
- Position: Don't Just "Fix it in Post": A Science of AI Must Study Training Dynamics
- Object Detection from Scratch with Deep Supervision
- Constrained Linear Data-feature Mapping for Image Classification
- DeepVisInterests: CNN-Ontology Prediction of Users Interests from Social Images
- Stochastic Adversarial Gradient Embedding for Active Domain Adaptation
- Vision Transformers Need Registers
- The Pace of Artificial Intelligence Innovations: Speed, Talent, and Trial-and-Error
- Copy-Move Forgery Classification via Unsupervised Domain Adaptation
- Fusion Recurrent Neural Network
- The Golden Ratio of Learning and Momentum
- Spatial Interpolation of Room Impulse Responses based on Deeper Physics-Informed Neural Networks with Residual Connections
- Evolution in Groups: A deeper look at synaptic cluster driven evolution of deep neural networks
- Accelerating CNN Training by Pruning Activation Gradients
- MCMC Guided CNN Training and Segmentation for Pancreas Extraction
- Time-Limited Toeplitz Operators on Abelian Groups: Applications in Information Theory and Subspace Approximation
- Survey of Dropout Methods for Deep Neural Networks
- Deep Selective Combinatorial Embedding and Consistency Regularization for Light Field Super-resolution
- APEX-Net: Automatic Plot Extractor Network
- Semantics, Representations and Grammars for Deep Learning
- Evaluating an Adaptive Multispectral Turret System for Autonomous Tracking Across Variable Illumination Conditions
- A Robust framework for sound event localization and detection on real recordings
- A Simple Framework for Contrastive Learning of Visual Representations
- SNIP: Single-shot Network Pruning based on Connection Sensitivity
- Optimal Gradient Quantization Condition for Communication-Efficient Distributed Training
- A Sensitivity Analysis of (and Practitioners' Guide to) Convolutional Neural Networks for Sentence Classification
- Deep Interactive Denoiser (DID) for X-Ray Computed Tomography
- Compression-aware Continual Learning using Singular Value Decomposition
- Segmentation of digital rock images using deep convolutional autoencoder networks
- ResIST: Layer-Wise Decomposition of ResNets for Distributed Training
- Recover and Identify: A Generative Dual Model for Cross-Resolution Person Re-Identification
- Improving Image Classification with Location Context
- Mutual Modality Trust with Lightweight Reconstruction Regularization for Fine-grained Tire Pattern Recognition
- A Sustainable Multi-modal Multi-layer Emotion-aware Service at the Edge
- Reasoning About Human-Object Interactions Through Dual Attention Networks
- Performance tuning for deep learning on a many-core processor (master thesis)
- Associatively Segmenting Instances and Semantics in Point Clouds
- Resolution Switchable Networks for Runtime Efficient Image Recognition
- American Sign Language fingerspelling recognition in the wild
- Transfer Learning Using Classification Layer Features of CNN
- The Loss Surfaces of Multilayer Networks
- TransNFCM: Translation-Based Neural Fashion Compatibility Modeling
- Learning Manipulation under Physics Constraints with Visual Perception
- Multi-Service Mobile Traffic Forecasting via Convolutional Long Short-Term Memories
- ESAI: Efficient Split Artificial Intelligence via Early Exiting Using Neural Architecture Search
- Deep inspection: an electrical distribution pole parts study via deep neural networks
- Improving robustness against common corruptions by covariate shift adaptation
- Node Classification on Graphs with Few-Shot Novel Labels via Meta Transformed Network Embedding
- Tuning Algorithms and Generators for Efficient Edge Inference
- Deep Image Orientation Angle Detection
- Towards High-Level Semantic Intelligence
- The neural architecture of language: Integrative modeling converges on predictive processing
- Data-driven emergence of convolutional structure in neural networks
- Neural model robustness for skill routing in large-scale conversational AI systems: A design choice exploration
- Knowledge Transfer via Distillation of Activation Boundaries Formed by Hidden Neurons
- Exploiting Kernel Sparsity and Entropy for Interpretable CNN Compression
- Basic Principles of Clustering Methods
- A predictor-corrector method for the training of deep neural networks
- Multi-Modal Object Re-Identification with Prompt-S6 and Semantic-Aware Knowledge Guidance
- WeMix: How to Better Utilize Data Augmentation
- Tell Me How to Ask Again: Question Data Augmentation with Controllable Rewriting in Continuous Space
- Accelerating Multi-Model Inference by Merging DNNs of Different Weights
- HABNet: Machine Learning, Remote Sensing Based Detection and Prediction of Harmful Algal Blooms
- A Theoretical Framework for Robustness of (Deep) Classifiers against Adversarial Examples
- Deep Tracking: Visual Tracking Using Deep Convolutional Networks
- Wavelet Integrated CNNs for Noise-Robust Image Classification
- Generalisation in humans and deep neural networks
- Physical world assistive signals for deep neural network classifiers -- neither defense nor attack
- Siamese Anchor Proposal Network for High-Speed Aerial Tracking
- GLAC Net: GLocal Attention Cascading Networks for Multi-image Cued Story Generation
- Learning from Multi-domain Artistic Images for Arbitrary Style Transfer
- Ablation Studies in Artificial Neural Networks
- Multilingual Image Description with Neural Sequence Models
- A CNN-RNN Architecture for Multi-Label Weather Recognition
- Joint Shape Representation and Classification for Detecting PDAC
- Adversarial Attack Type I: Cheat Classifiers by Significant Changes
- Extraction of Salient Sentences from Labelled Documents
- DeepMRSeg: A convolutional deep neural network for anatomy and abnormality segmentation on MR images
- MUSCLE: Strengthening Semi-Supervised Learning Via Concurrent Unsupervised Learning Using Mutual Information Maximization
- Iterative temporal differencing with random synaptic feedback weights support error backpropagation for deep learning
- Spectral Unsupervised Domain Adaptation for Visual Recognition
- P-ODN: Prototype based Open Deep Network for Open Set Recognition
- Sparse Vector Transmission: An Idea Whose Time Has Come
- Efficient Multi-Modal Embeddings from Structured Data
- Object Recognition from Short Videos for Robotic Perception
- Framing U-Net via Deep Convolutional Framelets: Application to Sparse-view CT
- Hypernetwork-Based Augmentation
- Post-Earthquake Assessment of Buildings Using Deep Learning
- Budgeted Training: Rethinking Deep Neural Network Training Under Resource Constraints
- Model-Based Domain Generalization
- Multi-loss ensemble deep learning for chest X-ray classification
- Robust Training in High Dimensions via Block Coordinate Geometric Median Descent
- CLAR: Contrastive Learning of Auditory Representations
- Statistical theory for image classification using deep convolutional neural networks with cross-entropy loss under the hierarchical max-pooling model
- Deep Transfer Learning for Automated Diagnosis of Skin Lesions from Photographs
- Deep Collective Learning: Learning Optimal Inputs and Weights Jointly in Deep Neural Networks
- Lagrangian Neural Networks
- SASL: Saliency-Adaptive Sparsity Learning for Neural Network Acceleration
- Fully Dynamic Inference with Deep Neural Networks
- An empirical investigation into the properties of standard word embeddings
- Validating predictions of burial mounds with field data: the promise and reality of machine learning
- Cascaded Structure Tensor Framework for Robust Identification of Heavily Occluded Baggage Items from X-ray Scans
- A Survey on Deep Geometry Learning: From a Representation Perspective
- Deep Anchored Convolutional Neural Networks
- DSCH-Loss: A Dynamic Semantic Channel Objective for Deep Semantic Hashing
- Single-shot Channel Pruning Based on Alternating Direction Method of Multipliers
- Mutually-aware Sub-Graphs Differentiable Architecture Search
- Axial-DeepLab: Stand-Alone Axial-Attention for Panoptic Segmentation
- AI Empowered Communication and Radar Modulation Recognition: A Survey
- Image-Based Geo-Localization Using Satellite Imagery
- Automated Numerical Stability Analysis of Deep Learning Operators
- Equivariant Q Learning in Spatial Action Spaces
- Guided Evolution for Neural Architecture Search
- Cuttlefish: A Lightweight Primitive for Adaptive Query Processing
- Deep Neural Network for Learning to Rank Query-Text Pairs
- Mind the Missing Split: Resolving Feature Heterogeneity in Swarm Learning with Random Forests
- On the Opportunities and Risks of Foundation Models
- Sparse Symplectically Integrated Neural Networks
- Image2Lego: Customized LEGO Set Generation from Images
- Why is AI hard and Physics simple?
- Deep Kinship Verification via Appearance-shape Joint Prediction and Adaptation-based Approach
- NeuroMAX: A High Throughput, Multi-Threaded, Log-Based Accelerator for Convolutional Neural Networks
- A Comprehensive Survey of Neural Architecture Search: Challenges and Solutions
- Same Predictions, Different Reasons: The Effect of Quantization on Model Explanations
- Real-time Reconstruction of Human Visual Perception from fMRI
- CLIP-Adapter: Better Vision-Language Models with Feature Adapters
- DEAL: Difficulty-aware Active Learning for Semantic Segmentation
- Guided Attention Network for Object Detection and Counting on Drones
- The Causal Loss: Driving Correlation to Imply Causation
- Text-to-image Synthesis via Symmetrical Distillation Networks
- Benchmarking deep learning models for Raman spectroscopy across open-source datasets
- SegSort: Segmentation by Discriminative Sorting of Segments
- Optimizing Resource Allocation for Geographically-Distributed Inference by Large Language Models
- LibContinual: A Comprehensive Library towards Realistic Continual Learning
- An Image is Worth 16x16 Words: Transformers for Image Recognition at Scale
- An Empirical Study of Machine Learning Robustness and Scalability for Imbalanced Tabular Clinical Data in Emergency and Critical Care
- Incremental Learning Using a Grow-and-Prune Paradigm with Efficient Neural Networks
- Detecting Medical Misinformation on Social Media Using Multimodal Deep Learning
- Granular Ball Guided Masking: Structure-aware Data Augmentation
- Programmable Optical Spectrum Shapers as Computing Primitives for Accelerating Convolutional Neural Networks
- Multi-Grained Text-Guided Image Fusion for Multi-Exposure and Multi-Focus Scenarios
- A Comprehensive Study of Bugs in Modern Distributed Deep Learning Systems
- Monocular Depth Estimation with Augmented Ordinal Depth Relationships
- Item Region-based Style Classification Network (IRSN): A Fashion Style Classifier Based on Domain Knowledge of Fashion Experts
- Bloom Filter Encoding for Machine Learning
- Statistically Significant Stopping of Neural Network Training
- Image-specific Convolutional Kernel Modulation for Single Image Super-resolution
- Handling Inter-Annotator Agreement for Automated Skin Lesion Segmentation
- The Seismic Wavefield Common Task Framework
- Intrinsic dimension of data representations in deep neural networks
- Combating Adversarial Misspellings with Robust Word Recognition
- DK-STN: A Domain Knowledge Embedded Spatio-Temporal Network Model for MJO Forecast
- A Convolutional Neural Deferred Shader for Physics Based Rendering
- Deep Kernel Learning
- Predicting Human Trajectories by Learning and Matching Patterns
- Brain-Gen: Towards Interpreting Neural Signals for Stimulus Reconstruction Using Transformers and Latent Diffusion Models
- Adversarial Robustness in Zero-Shot Learning:An Empirical Study on Class and Concept-Level Vulnerabilities
- The Interaction Bottleneck of Deep Neural Networks: Discovery, Proof, and Modulation
- Towards Ancient Plant Seed Classification: A Benchmark Dataset and Baseline Model
- Adversarial Robustness of Vision in Open Foundation Models
- Semi-Supervised Online Learning on the Edge by Transforming Knowledge from Teacher Models
- KOSS: Kalman-Optimal Selective State Spaces for Long-Term Sequence Modeling
- SARMAE: Masked Autoencoder for SAR Representation Learning
- Batch Normalization-Free Fully Integer Quantized Neural Networks via Progressive Tandem Learning
- Hypernetworks That Evolve Themselves
- AI Needs Physics More Than Physics Needs AI
- Higher-Order LaSDI: Reduced Order Modeling with Multiple Time Derivatives
- In Pursuit of Pixel Supervision for Visual Pre-training
- Stylized Synthetic Augmentation further improves Corruption Robustness
- From Risk to Resilience: Towards Assessing and Mitigating the Risk of Data Reconstruction Attacks in Federated Learning
- Image Complexity-Aware Adaptive Retrieval for Efficient Vision-Language Models
- An updated efficient galaxy morphology classification model based on ConvNeXt encoding with UMAP dimensionality reduction
- TrajSyn: Privacy-Preserving Dataset Distillation from Federated Model Trajectories for Server-Side Adversarial Training
- Which Coauthor Should I Nominate in My 99 ICLR Submissions? A Mathematical Analysis of the ICLR 2026 Reciprocal Reviewer Nomination Policy
- Prototypical Learning Guided Context-Aware Segmentation Network for Few-Shot Anomaly Detection
- Mimicking Human Visual Development for Learning Robust Image Representations
- DriverGaze360: OmniDirectional Driver Attention with Object-Level Guidance
- LFFD: A Light and Fast Face Detector for Edge Devices
- Optimizing the Adversarial Perturbation with a Momentum-based Adaptive Matrix
- On Improving Deep Active Learning with Formal Verification
- Multiple weak biases support adaptive choices without prior experience: a self-supervised strategy
- Unbiased Mean Teacher for Cross-domain Object Detection
- Learning Deep Bilinear Transformation for Fine-grained Image Representation
- Bilinear CNNs for Fine-grained Visual Recognition
- Efficient Dense Modules of Asymmetric Convolution for Real-Time Semantic Segmentation
- Gradient descent aligns the layers of deep linear networks
- Practical Implementation of Memristor-Based Threshold Logic Gates
- Asymptotic Soft Filter Pruning for Deep Convolutional Neural Networks
- Dropout Neural Network Training Viewed from a Percolation Perspective
- Evaluating Singular Value Thresholds for DNN Weight Matrices based on Random Matrix Theory
- Multiclass Graph-Based Large Margin Classifiers: Unified Approach for Support Vectors and Neural Networks
- Explanatory models in neuroscience: Part 1 -- taking mechanistic abstraction seriously
- Effective Fine-Tuning with Eigenvector Centrality Based Pruning
- Generative Spatiotemporal Data Augmentation
- Advancing Cache-Based Few-Shot Classification via Patch-Driven Relational Gated Graph Attention
- Machine learning methods for subpixel trajectory reconstruction in discretized position detectors
- Efficient Classification of Very Large Images with Tiny Objects
- View Space: Learning Representation across Arbitrary Graphs
- DREAM-B3P: Dual-Stream Transformer Network Enhanced by Feedback Diffusion Model for Blood-Brain Barrier Penetrating Peptide Prediction
- Where Classification Fails, Interpretation Rises
- Understanding Deep Learning Techniques for Image Segmentation
- FaaT: A Transparent Auto-Scaling Cache for Serverless Applications
- Back to the Baseline: Examining Baseline Effects on Explainability Metrics
- Application of a semantic segmentation convolutional neural network for accurate automatic detection and mapping of solar photovoltaic arrays in aerial imagery
- A Review of Machine Learning Applications in Fuzzing
- Visual Search Asymmetry: Deep Nets and Humans Share Similar Inherent Biases
- Quantum Algorithms for Unsupervised Machine Learning and Neural Networks
- Automated Detection of Equine Facial Action Units
- Poker-CNN: A Pattern Learning Strategy for Making Draws and Bets in Poker Games
- Take a Peek: Efficient Encoder Adaptation for Few-Shot Semantic Segmentation via LoRA
- Attending to Discriminative Certainty for Domain Adaptation
- An Overview and Case Study of the Clinical AI Model Development Life Cycle for Healthcare Systems
- SGNet: A Super-class Guided Network for Image Classification and Object Detection
- Federated Domain Generalization with Latent Space Inversion
- Learning to Optimize Tensor Programs
- ZenSVI: An open-source software for the integrated acquisition, processing and analysis of street view imagery towards scalable urban science
- Dynamic Efficient Adversarial Training Guided by Gradient Magnitude
- Learning Quantized Neural Nets by Coarse Gradient Method for Non-linear Classification
- Hands-on Evaluation of Visual Transformers for Object Recognition and Detection
- DirectSwap: Mask-Free Cross-Identity Training and Benchmarking for Expression-Consistent Video Head Swapping
- Microscopic Vehicle Trajectories from Heterogeneous and Area-Based Traffic
- Causality-inspired Single-source Domain Generalization for Medical Image Segmentation
- A Review of Robot Learning for Manipulation: Challenges, Representations, and Algorithms
- Including Image-based Perception in Disturbance Observer for Warehouse Drones
- An Approach for Detection of Entities in Dynamic Media Contents
- Chopper: A Multi-Level GPU Characterization Tool & Derived Insights Into LLM Training Inefficiency
- PR-CapsNet: Pseudo-Riemannian Capsule Network with Adaptive Curvature Routing for Graph Learning
- LayerPipe2: Multistage Pipelining and Weight Recompute via Improved Exponential Moving Average for Training Neural Networks
- On the Heavy-Tailed Theory of Stochastic Gradient Descent for Deep Neural Networks
- Fully Convolutional Neural Networks for Dynamic Object Detection in Grid Maps (Masters Thesis)
- Atlas: A Dataset and Benchmark for E-commerce Clothing Product\n Categorization
- Natural-Logarithm-Rectified Activation Function in Convolutional Neural Networks
- PCGAN-CHAR: Progressively Trained Classifier Generative Adversarial Networks for Classification of Noisy Handwritten Bangla Characters
- A Closer Look at Spatiotemporal Convolutions for Action Recognition
- Tackling Graphical NLP problems with Graph Recurrent Networks
- HOLE: Homological Observation of Latent Embeddings for Neural Network Interpretability
- PolSAR Image Classification Based on Dilated Convolution and Pixel-Refining Parallel Mapping network in the Complex Domain
- LVIS: A Dataset for Large Vocabulary Instance Segmentation
- Revisiting Latent-Space Interpolation via a Quantitative Evaluation Framework
- Barrier-Free Large-Scale Sparse Tensor Accelerator (BARISTA) For\n Convolutional Neural Networks
- Explaining Deep Neural Networks
- Life is not black and white -- Combining Semi-Supervised Learning with fuzzy labels
- Amulet: Fast TEE-Shielded Inference for On-Device Model Protection
- GlimmerNet: A Lightweight Grouped Dilated Depthwise Convolutions for UAV-Based Emergency Monitoring
- Enhanced Chest Disease Classification Using an Improved CheXNet Framework with EfficientNetV2-M and Optimization-Driven Learning
- Integrating Multi-scale and Multi-filtration Topological Features for Medical Image Classification
- Towards Robust Pseudo-Label Learning in Semantic Segmentation: An Encoding Perspective
- From Forecast to Action: A Deep Learning Model for Predicting Power Outages During Tropical Cyclones
- Hierarchical Deep Learning for Diatom Image Classification: A Multi-Level Taxonomic Approach
- ECA-Net: Efficient Channel Attention for Deep Convolutional Neural Networks
- A Perception CNN for Facial Expression Recognition
- CN-CELEB: a challenging Chinese speaker recognition dataset
- Analysis of Social Media Data using Multimodal Deep Learning for Disaster Response
- Cross-Image Region Mining with Region Prototypical Network for Weakly Supervised Segmentation
- Tensor object classification via multilinear discriminant analysis network
- Novel Deep Learning Architectures for Classification and Segmentation of Brain Tumors from MRI Images
- Statistical physics for artificial neural networks
- Towards Hardware-Agnostic Gaze-Trackers
- SPOOF: Simple Pixel Operations for Out-of-Distribution Fooling
- A Comparative Study on Synthetic Facial Data Generation Techniques for Face Recognition
- Achieving Approximate Symmetry Is Exponentially Easier than Exact Symmetry
- AI & Human Co-Improvement for Safer Co-Superintelligence
- Affine Non-negative Collaborative Representation Based Pattern Classification
- CNN on `Top': In Search of Scalable & Lightweight Image-based Jet Taggers
- TEA: Temporal Excitation and Aggregation for Action Recognition
- On the Effect of Regularization on Nonparametric Mean-Variance Regression
- A Survey of Convolutional Neural Networks: Analysis, Applications, and Prospects
- Generative Recursive Reasoning
- ConvSequential-SLAM: A Sequence-based, Training-less Visual Place Recognition Technique for Changing Environments
- A Comparative Review of Recent Few-Shot Object Detection Algorithms
- Human-Level Control without Server-Grade Hardware
- TactileSGNet: A Spiking Graph Neural Network for Event-based Tactile Object Recognition
- An all-optical convolutional neural network for image identification
- A Comprehensive Survey of Machine Learning Applied to Radar Signal Processing
- A Cascaded Zoom-In Network for Patterned Fabric Defect Detection
- Rethinking Decoupled Knowledge Distillation: A Predictive Distribution Perspective
- QoSDiff: An Implicit Topological Embedding Learning Framework Leveraging Denoising Diffusion and Adversarial Attention for Robust QoS Prediction
- CASTLE: Regularization via Auxiliary Causal Graph Discovery
- RRPN++: Guidance Towards More Accurate Scene Text Detection
- Performance Evaluation of Transfer Learning Based Medical Image Classification Techniques for Disease Detection
- Bayes-DIC Net: Estimating Digital Image Correlation Uncertainty with Bayesian Neural Networks
- CoDA: From Text-to-Image Diffusion Models to Training-Free Dataset Distillation
- Parameter efficient hybrid spiking-quantum convolutional neural network with surrogate gradient and quantum data-reupload
- Deep Unfolding: Recent Developments, Theory, and Design Guidelines
- On the Binding Problem in Artificial Neural Networks
- Multi-Scale Visual Prompting for Lightweight Small-Image Classification
- What does LIME really see in images?
- Compressed Video Action Recognition with Refined Motion Vector
- Using Deep Learning and Machine Learning to Detect Epileptic Seizure with Electroencephalography (EEG) Data
- Group-Structured Adversarial Training
- Detecting Electric Devices in 3D Images of Bags
- Benchmarking machine learning models for multi-class state recognition in double quantum dot data
- Leveraging Large-Scale Pretrained Spatial-Spectral Priors for General Zero-Shot Pansharpening
- OmniPerson: Unified Identity-Preserving Pedestrian Generation
- Associative Memory using Attribute-Specific Neuron Groups-1: Learning between Multiple Cue Balls
- Breast Cell Segmentation Under Extreme Data Constraints: Quantum Enhancement Meets Adaptive Loss Stabilization
- The brain-AI convergence: Predictive and generative world models for general-purpose computation
- TGDD: Trajectory Guided Dataset Distillation with Balanced Distribution
- Dependency Aware Filter Pruning
- Classification of Radio Signals and HF Transmission Modes with Deep Learning
- Equilibrium Propagation Without Limits
- TPCNet: Triple physical constraints for Low-light Image Enhancement
- Large Scale Fine-Grained Categorization and Domain-Specific Transfer Learning
- Multimodal Mixture-of-Experts for ISAC in Low-Altitude Wireless Networks
- Mid-Level Visual Representations Improve Generalization and Sample Efficiency for Learning Visuomotor Policies
- Neural Networks for Predicting Permeability Tensors of 2D Porous Media: Comparison of Convolution- and Transformer-based Architectures
- Deep Learning for Generic Object Detection: A Survey
- Deep Learning in Neural Networks: An Overview
- A systematic study of the class imbalance problem in convolutional neural networks
- Masked Autoencoders Are Scalable Vision Learners
- SRM : A Style-based Recalibration Module for Convolutional Neural Networks
- Compact representations of convolutional neural networks via weight pruning and quantization
- CrossedWires: A Dataset of Syntactically Equivalent but Semantically Disparate Deep Learning Models
- Non-Parametric Neural Style Transfer
- Differentiable Weightless Controllers: Learning Logic Circuits for Continuous Control
- An Introduction to Convolutional Neural Networks
- Directed evolution algorithm drives neural prediction
- Spatiotemporal Satellite Image Downscaling with Transfer Encoders and Autoregressive Generative Models
- Empowering Things with Intelligence: A Survey of the Progress, Challenges, and Opportunities in Artificial Intelligence of Things
- Long-term Temporal Convolutions for Action Recognition
- Emotion Recognition from Speech
- IGen: Scalable Data Generation for Robot Learning from Open-World Images
- Controllable 3D Object Generation with Single Image Prompt
- Scaling Laws for Neural Language Models
- Novelty Detection in MultiClass Scenarios with Incomplete Set of Class Labels
- All-Weather Object Recognition Using Radar and Infrared Sensing
- Dual-Projection Fusion for Accurate Upright Panorama Generation in Robotic Vision
- Designing a Micro-Benchmark Suite to Evaluate gRPC for TensorFlow: Early Experiences
- ODE guided Neural Data Augmentation Techniques for Time Series Data and its Benefits on Robustness
- CACARA: Cross-Modal Alignment Leveraging a Text-Centric Approach for Cost-Effective Multimodal and Multilingual Learning
- Structured Context Learning for Generic Event Boundary Detection
- Fixing the train-test resolution discrepancy
- N-ImageNet: Towards Robust, Fine-Grained Object Recognition with Event Cameras
- BioArc: Discovering Optimal Neural Architectures for Biological Foundation Models
- "Why the face?": Exploring Robot Error Detection Using Instrumented Bystander Reactions
- SelfVIO: Self-supervised deep monocular Visual–Inertial Odometry and depth estimation
- First Steps towards Machine Learning for Prediction and Pre-Correction in Direct Laser Writing
- GhostNet: More Features from Cheap Operations
- Explaining Deep Learning Models for Structured Data using Layer-Wise Relevance Propagation
- Action Recognition with Kernel-based Graph Convolutional Networks
- A Unified Framework for Multi-View Multi-Class Object Pose Estimation
- Bharat Scene Text: A Novel Comprehensive Dataset and Benchmark for Indian Language Scene Text Understanding
- Local and Global Context-and-Object-part-Aware Superpixel-based Data Augmentation for Deep Visual Recognition
- CyCNN: A Rotation Invariant CNN using Polar Mapping and Cylindrical Convolution Layers
- Learning Deep Structure-Preserving Image-Text Embeddings
- CycleCluster: Modernising Clustering Regularisation for Deep Semi-Supervised Classification
- On-the-Job Learning with Bayesian Decision Theory
- CNN-Based Framework for Pedestrian Age and Gender Classification Using Far-View Surveillance in Mixed-Traffic Intersections
- Hybrid Context-Fusion Attention (CFA) U-Net and Clustering for Robust Seismic Horizon Interpretation
- DAONet-YOLOv8: An Occlusion-Aware Dual-Attention Network for Tea Leaf Pest and Disease Detection
- Adversarial Training Towards Robust Multimedia Recommender System
- Adversarial AutoMixup
- Active Learning for GCN-based Action Recognition
- tfShearlab: The TensorFlow Digital Shearlet Transform for Deep Learning
- A Variational View on Bootstrap Ensembles as Bayesian Inference
- Neurodevelopmental Age Estimation of Infants Using a 3D-Convolutional Neural Network Model based on Fusion MRI Sequences
- Neural Networks, Hypersurfaces, and Radon Transforms
- Photo-Realistic Video Prediction on Natural Videos of Largely Changing Frames
- A multi-task convolutional neural network for mega-city analysis using very high resolution satellite imagery and geospatial data
- Provably Powerful Graph Networks
- Tactile-Based Insertion for Dense Box-Packing
- Decoupled Dynamic Filter Networks
- STEP: Segmenting and Tracking Every Pixel
- Underexposed Image Correction via Hybrid Priors Navigated Deep Propagation
- FOSNet: An End-to-End Trainable Deep Neural Network for Scene Recognition
- Fundamentals of Regression
- Optimal Conversion of Conventional Artificial Neural Networks to Spiking Neural Networks
- BiSeNet V2: Bilateral Network with Guided Aggregation for Real-Time Semantic Segmentation
- Mean-Field Model for Two-Layer Neural Networks Trained with Consensus-Based Optimization
- Differentiable Physics-Neural Models enable Learning of Non-Markovian Closures for Accelerated Coarse-Grained Physics Simulations
- A Physics-Informed U-net-LSTM Network for Data-Driven Seismic Response Modeling of Structures
- MorphingDB: A Task-Centric AI-Native DBMS for Model Management and Inference
- Physics-informed neural networks method in high-dimensional integrable systems
- ShuffleNet: An Extremely Efficient Convolutional Neural Network for Mobile Devices
- Generalized Out-of-Distribution Detection: A Survey
- Privacy-Preserving Self-Taught Federated Learning for Heterogeneous Data
- Improving the Generalization of End-to-End Driving through Procedural Generation
- WildDeepfake: A Challenging Real-World Dataset for Deepfake Detection
- Comprehensive Graph-conditional Similarity Preserving Network for Unsupervised Cross-modal Hashing
- Deep Learning Assisted Calibrated Beam Training for Millimeter-Wave Communication Systems
- Guaranteed Optimal Compositional Explanations for Neurons
- Open Vocabulary Compositional Explanations for Neuron Alignment
- CarBench: A Comprehensive Benchmark for Neural Surrogates on High-Fidelity 3D Car Aerodynamics
- On a Sparse Shortcut Topology of Artificial Neural Networks
- Pre-train to Gain: Robust Learning Without Clean Labels
- NNGPT: Rethinking AutoML with Large Language Models
- Back to the Feature: Explaining Video Classifiers with Video Counterfactual Explanations
- Advancing Image Classification with Discrete Diffusion Classification Modeling
- Exo2EgoSyn: Unlocking Foundation Video Generation Models for Exocentric-to-Egocentric Video Synthesis
- Federated Learning in Mobile Edge Networks: A Comprehensive Survey
- Triplet-Based Deep Hashing Network for Cross-Modal Retrieval
- Temporally Distributed Networks for Fast Video Semantic Segmentation
- gradSLAM: Automagically differentiable SLAM
- C3 Framework: An Open-source PyTorch Code for Crowd Counting
- ModHiFi: Identifying High Fidelity predictive components for Model Modification
- IDSplat: Instance-Decomposed 3D Gaussian Splatting for Driving Scenes
- Towards Deep Learning Models Resistant to Adversarial Attacks
- Visual Imitation Made Easy
- Experimental insights into data augmentation techniques for deep learning-based multimode fiber imaging: limitations and success
- Prototype-supervised Adversarial Network for Targeted Attack of Deep Hashing
- Combating Ambiguity for Hash-code Learning in Medical Instance Retrieval
- Dynamic Granularity Matters: Rethinking Vision Transformers Beyond Fixed Patch Splitting
- Deep fusion of multi-view and multimodal representation of ALS point cloud for 3D terrain scene recognition
- Towards Biologically Plausible Convolutional Networks
- Multimodal Real-Time Anomaly Detection and Industrial Applications
- Phase-Aligned RoPE for Mixed-Resolution Diffusion Transformer
- LookSharp: Attention Entropy Minimization for Test-Time Adaptation
- FHE-Agent: Automating CKKS Configuration for Practical Encrypted Inference via an LLM-Guided Agentic Framework
- Initializing ReLU networks in an expressive subspace of weights
- Cross-individual Recognition of Emotions by a Dynamic Entropy based on Pattern Learning with EEG features
- A Climatology of Quasi-Linear Convective Systems and Their Hazards in the United States
- Enhanced Center Coding for Cell Detection with Convolutional Neural Networks
- RAISECity: A Multimodal Agent Framework for Reality-Aligned 3D World Generation at City-Scale
- FAST: Topology-Aware Frequency-Domain Distribution Matching for Coreset Selection
- Using MLIR Transform to Design Sliced Convolution Algorithm
- Decoupled Audio-Visual Dataset Distillation
- APRIL: Annotations for Policy evaluation with Reliable Inference from LLMs
- The Rapid Growth of AI Foundation Model Usage in Science
- Responses to Critiques on Machine Learning of Criminality Perceptions (Addendum of arXiv:1611.04135)
- Neural Architecture Search without Training
- Separation of time scales and direct computation of weights in deep neural networks
- LassoLayer: Nonlinear Feature Selection by Switching One-to-one Links
- Multiple Code Hashing for Efficient Image Retrieval
- PP-YOLO: An Effective and Efficient Implementation of Object Detector
- Spatial-Temporal Dynamic Graph Attention Networks for Ride-hailing Demand Prediction
- Layer-wise Weight Selection for Power-Efficient Neural Network Acceleration
- An End-to-End Approach to Automatic Speech Assessment for Cantonese-speaking People with Aphasia
- Enhancing Adversarial Transferability through Block Stretch and Shrink
- Neural Collapse Under MSE Loss: Proximity to and Dynamics on the Central Path
- Feasibility of Embodied Dynamics Based Bayesian Learning for Continuous Pursuit Motion Control of Assistive Mobile Robots in the Built Environment
- MorphSeek: Fine-grained Latent Representation-Level Policy Optimization for Deformable Image Registration
- Using Topological Framework for the Design of Activation Function and Model Pruning in Deep Neural Networks
- Automatic Recognition of Coal and Gangue based on Convolution Neural Network
- A Machine Learning-Driven Solution for Denoising Inertial Confinement Fusion Images
- Evolution Strategies at the Hyperscale
- Compositional Explanations of Neurons
- Learning from Noisy Labels with Distillation
- Spatial-Temporal Transformer Networks for Traffic Flow Forecasting
- Learning without Forgetting
- UniFormer: Unifying Convolution and Self-Attention for Visual Recognition
- FVQA: Fact-Based Visual Question Answering
- Network Pruning for Low-Rank Binary Indexing
- Breast Tumor Classification and Segmentation using Convolutional Neural Networks
- CathAI: Fully Automated Interpretation of Coronary Angiograms Using Neural Networks
- GLOBE: Accurate and Generalizable PDE Surrogates using Domain-Inspired Architectures and Equivariances
- A Review of Machine Learning for Cavitation Intensity Recognition in Complex Industrial Systems
- Learning to Navigate Using Mid-Level Visual Priors
- What Your Features Reveal: Data-Efficient Black-Box Feature Inversion Attack for Split DNNs
- Med3D: Transfer Learning for 3D Medical Image Analysis
- Task-driven Semantic Coding via Reinforcement Learning
- Incidental Scene Text Understanding: Recent Progresses on ICDAR 2015 Robust Reading Competition Challenge 4
- Syn2Real: Forgery Classification via Unsupervised Domain Adaptation
- A Novel Perspective to Zero-shot Learning: Towards an Alignment of Manifold Structures via Semantic Feature Expansion
- Clustered Object Detection in Aerial Images
- Learning to See Through a Baby's Eyes: Early Visual Diets Enable Robust Visual Intelligence in Humans and Machines
- Sigil: Server-Enforced Watermarking in U-Shaped Split Federated Learning via Gradient Injection
- VisEvent: Reliable Object Tracking via Collaboration of Frame and Event Flows
- Deep Discriminative Representation Learning with Attention Map for Scene Classification
- SQWA: Stochastic Quantized Weight Averaging for Improving the Generalization Capability of Low-Precision Deep Neural Networks
- XNOR-Net++: Improved Binary Neural Networks
- Towards Optimal Structured CNN Pruning via Generative Adversarial Learning
- DCNNs: A Transfer Learning comparison of Full Weapon Family threat detection for Dual-Energy X-Ray Baggage Imagery
- Convolutional Neural Networks for Large-Scale Remote-Sensing Image Classification
- EVAA—Exchange Vanishing Adversarial Attack on LiDAR Point Clouds in Autonomous Vehicles
- Cross-Learning from Scarce Data via Multi-Task Constrained Optimization
- Hamiltonian Neural Networks
- Learning proofs for the classification of nilpotent semigroups
- Region-Point Joint Representation for Effective Trajectory Similarity Learning
- Surgical Visual Domain Adaptation: Results from the MICCAI 2020 SurgVisDom Challenge
- Compositional Generalization for Primitive Substitutions
- Highly Scalable Deep Learning Training System with Mixed-Precision: Training ImageNet in Four Minutes
- Sequence to Sequence Learning with Neural Networks
- IntPhys: A Framework and Benchmark for Visual Intuitive Physics Reasoning
- BinaryConnect: Training Deep Neural Networks with binary weights during propagations
- Questions to Guide the Future of Artificial Intelligence Research
- Recommendation or Discrimination?: Quantifying Distribution Parity in Information Retrieval Systems
- Generic decoding of seen and imagined objects using hierarchical visual features
- Information Theory-Guided Heuristic Progressive Multi-View Coding
- To Click or Not To Click: Automatic Selection of Beautiful Thumbnails from Videos
- Learning a Deep Embedding Model for Zero-Shot Learning
- Multi-Interactive Attention Network for Fine-grained Feature Learning in CTR Prediction
- TinyCNN: A Tiny Modular CNN Accelerator for Embedded FPGA
- MineGAN: effective knowledge transfer from GANs to target domains with few images
- A Gaussian Process perspective on Convolutional Neural Networks
- Alpha-Integration Pooling for Convolutional Neural Networks
- Depth Map Prediction from a Single Image using a Multi-Scale Deep Network
- Meta Learning with Differentiable Closed-form Solver for Fast Video Object Segmentation
- LaSOT: A High-quality Large-scale Single Object Tracking Benchmark
- Learning to Hash with Graph Neural Networks for Recommender Systems
- LAYA: Layer-wise Attention Aggregation for Interpretable Depth-Aware Neural Networks
- Sparsity-Control Ternary Weight Networks
- Quaternion Convolutional Neural Networks
- A Survey of Deep Reinforcement Learning in Video Games
- 1st Place Solution for Waymo Open Dataset Challenge -- 3D Detection and Domain Adaptation
- Understanding the robustness of deep neural network classifiers for breast cancer screening
- LEA-Net: Layer-wise External Attention Network for Efficient Color Anomaly Detection
- Calibrated Adversarial Training
- Separation and Concentration in Deep Networks
- Beyond Dents and Scratches: Logical Constraints in Unsupervised Anomaly Detection and Localization
- Be Your Own Teacher: Improve the Performance of Convolutional Neural Networks via Self Distillation
- Balanced Symmetric Cross Entropy for Large Scale Imbalanced and Noisy Data
- DeepTest: Automated Testing of Deep-Neural-Network-driven Autonomous Cars
- Credit scoring using neural networks and SURE posterior probability calibration
- Glance-and-Gaze Vision Transformer
- Memory In Memory: A Predictive Neural Network for Learning Higher-Order Non-Stationarity from Spatiotemporal Dynamics
- Efficient Transfer Bayesian Optimization with Auxiliary Information
- Discovery and Separation of Features for Invariant Representation Learning
- Accelerating Robustness Verification of Deep Neural Networks Guided by Target Labels
- Batch Normalization Preconditioning for Neural Network Training
- Learning Accurate, Comfortable and Human-like Driving
- End-to-End Wireframe Parsing
- Walsh-Hadamard Variational Inference for Bayesian Deep Learning
- Shifted Chunk Transformer for Spatio-Temporal Representational Learning
- Tackling Over-Smoothing for General Graph Convolutional Networks
- Model Patching: Closing the Subgroup Performance Gap with Data Augmentation
- Multi-level Texture Encoding and Representation (MuLTER) based on Deep Neural Networks
- Dense neural networks as sparse graphs and the lightning initialization
- Curriculum Learning: A Survey
- A Comprehensive Overhaul of Feature Distillation
- Can deep learning help you find the perfect match?
- Estimating or Propagating Gradients Through Stochastic Neurons
- HexCNN: A Framework for Native Hexagonal Convolutional Neural Networks
- Differentiable Rendering: A Survey
- Exemplar Loss for Siamese Network in Visual Tracking
- Extremal Contours: Gradient-driven contours for compact visual attribution
- Effective Regularization Through Loss-Function Metalearning
- Deep Image Prior
- Quantile regression with deep ReLU Networks: Estimators and minimax rates
- A study on using image based machine learning methods to develop the surrogate models of stamp forming simulations
- FLClear: Visually Verifiable Multi-Client Watermarking for Federated Learning
- Blow: a single-scale hyperconditioned flow for non-parallel raw-audio voice conversion
- Network Moments: Extensions and Sparse-Smooth Attacks
- Global Image Sentiment Transfer
- Deep Model-Based Reinforcement Learning for High-Dimensional Problems, a Survey
- Deep Learning Based Defect Detection for Solder Joints on Industrial X-Ray Circuit Board Images
- A Multicollinearity-Aware Signal-Processing Framework for Cross-β Identification via X-ray Scattering of Alzheimer's Tissue
- Multiscale Vision Transformers
- Machine Learning Framework for Efficient Prediction of Quantum Wasserstein Distance
- Learning Straight Flows: Variational Flow Matching for Efficient Generation
- Deep Spatial Pyramid: The Devil is Once Again in the Details
- Generalizing to unseen domains via distribution matching
- Hi-DREAM: Brain Inspired Hierarchical Diffusion for fMRI Reconstruction via ROI Encoder and visuAl Mapping
- Compiling to linear neurons
- The modified Physics-Informed Hybrid Parallel Kolmogorov--Arnold and Multilayer Perceptron Architecture with domain decomposition
- LEMUR: Large scale End-to-end MUltimodal Recommendation
- Physics-informed Machine Learning for Static Friction Modeling in Robotic Manipulators Based on Kolmogorov-Arnold Networks
- AdaptViG: Adaptive Vision GNN with Exponential Decay Gating
- AudioNet: Supervised Deep Hashing for Retrieval of Similar Audio Events
- Classification of motor faults based on transmission coefficient and reflection coefficient of omni-directional antenna using DCNN
- SliderEdit: Continuous Image Editing with Fine-Grained Instruction Control
- Why Should the Server Do It All?: A Scalable, Versatile, and Model-Agnostic Framework for Server-Light DNN Inference over Massively Distributed Clients via Training-Free Intermediate Feature Compression
- Multi-task Learning with Attention for End-to-end Autonomous Driving
- Learning Memory-guided Normality for Anomaly Detection
- Self-Guided Adaptation: Progressive Representation Alignment for Domain Adaptive Object Detection
- Unsupervised Domain Expansion from Multiple Sources
- Robust Policies via Mid-Level Visual Representations: An Experimental Study in Manipulation and Navigation
- RGBD Based Dimensional Decomposition Residual Network for 3D Semantic Scene Completion
- Chameleon: Adaptive Code Optimization for Expedited Deep Neural Network Compilation
- CutMix: Regularization Strategy to Train Strong Classifiers with Localizable Features
- A Comprehensive Study of Deep Video Action Recognition
- Convolutional neural networks decode visual stimulus positions from local field potentials on the mouse cortex
- C3AE: Exploring the Limits of Compact Model for Age Estimation
- State-Aware Tracker for Real-Time Video Object Segmentation
- PP-LCNet: A Lightweight CPU Convolutional Neural Network
- ProbSelect: Stochastic Client Selection for GPU-Accelerated Compute Devices in the 3D Continuum
- HipKittens: Fast and Furious AMD Kernels
- Advancing credibility and transparency in brain-to-image reconstruction research: Reanalysis of Koide-Majima, Nishimoto, and Majima (Neural Networks, 2024)
- Privacy Beyond Pixels: Latent Anonymization for Privacy-Preserving Video Understanding
- Quality-Aware Network for Human Parsing
- Integrating Large Circular Kernels into CNNs through Neural Architecture Search
- Under the Skin of Foundation NFT Auctions
- Deep Semantic Hashing with Generative Adversarial Networks
- Video Playback Rate Perception for Self-supervisedSpatio-Temporal Representation Learning
- Sample Prior Guided Robust Model Learning to Suppress Noisy Labels
- Edge of chaos as a guiding principle for modern neural network training
- Pose-Guided Multi-Granularity Attention Network for Text-Based Person Search
- AuxBlocks: Defense Adversarial Example via Auxiliary Blocks
- A Hybrid Autoencoder-Transformer Model for Robust Day-Ahead Electricity Price Forecasting under Extreme Conditions
- Minimum Width of Deep Narrow Networks for Universal Approximation
- Learning Performance Optimization for Edge AI System with Time and Energy Constraints
- Refactoring Neural Networks for Verification
- A Survey of Machine Learning Methods and Challenges for Windows Malware Classification
- Inter-Image Communication for Weakly Supervised Localization
- Preparation of Fractal-Inspired Computational Architectures for Advanced Large Language Model Analysis
- Anti-aliasing Deep Image Classifiers using Novel Depth Adaptive Blurring and Activation Function
- TopoAct: Visually Exploring the Shape of Activations in Deep Learning
- CBIR using Pre-Trained Neural Networks
- The Low-Rank Simplicity Bias in Deep Networks
- Rethinking Parameter Sharing as Graph Coloring for Structured Compression
- Real-Time Face and Landmark Localization for Eyeblink Detection
- Tip-Adapter: Training-free CLIP-Adapter for Better Vision-Language Modeling
- Escaping the Big Data Paradigm with Compact Transformers
- A Visual Perception-Based Tunable Framework and Evaluation Benchmark for H.265/HEVC ROI Encryption
- QDCNN: Quantum Dilated Convolutional Neural Network
- Qu-ANTI-zation: Exploiting Quantization Artifacts for Achieving Adversarial Outcomes
- Enhancing Multimodal Misinformation Detection by Replaying the Whole Story from Image Modality Perspective
- One-Shot Knowledge Transfer for Scalable Person Re-Identification
- Radio AGN feedback sustains quiescence only in a minority of massive galaxies
- Global Multiple Extraction Network for Low-Resolution Facial Expression Recognition
- Adversarial Attacks Beyond the Image Space
- Deep Image: Scaling up Image Recognition
- Toward Better Generalization in Few-Shot Learning through the Meta-Component Combination
- Global Context Networks
- Context-aware Learned Mesh-based Simulation via Trajectory-Level Meta-Learning
- NeuroFlex: Column-Exact ANN-SNN Co-Execution Accelerator with Cost-Guided Scheduling
- Beta Distribution Learning for Reliable Roadway Crash Risk Assessment
- Nowcast3D: Reliable precipitation nowcasting via gray-box learning
- Comparative Study of CNN Architectures for Binary Classification of Horses and Motorcycles in the VOC 2008 Dataset
- Accelerating scientific discovery with the common task framework
- Meta-World: A Benchmark and Evaluation for Multi-Task and Meta Reinforcement Learning
- Distribution-Aware Tensor Decomposition for Compression of Convolutional Neural Networks
- Extended Physics Informed Neural Network for Hyperbolic Two-Phase Flow in Porous Media
- Depth-induced NTK: Bridging Over-parameterized Neural Networks and Deep Neural Kernels
- SAAIPAA: Optimizing aspect-angles-invariant physical adversarial attacks on SAR target recognition models
- Joint Optimization of DNN Model Caching and Request Routing in Mobile Edge Computing
- DKN: Deep Knowledge-Aware Network for News Recommendation
- LeViT: a Vision Transformer in ConvNet's Clothing for Faster Inference
- A novel method for identifying the deep neural network model with the Serial Number
- On Complex Valued Convolutional Neural Networks
- Two at Once: Enhancing Learning and Generalization Capacities via IBN-Net
- SinReQ: Generalized Sinusoidal Regularization for Low-Bitwidth Deep Quantized Training
- Detecting cutaneous basal cell carcinomas in ultra-high resolution and weakly labelled histopathological images
- Deeper and Wider Siamese Networks for Real-Time Visual Tracking
- SPEC2: SPECtral SParsE CNN Accelerator on FPGAs
- An open access repository of images on plant health to enable the development of mobile disease diagnostics
- Learning with less: label-efficient land cover classification at very high spatial resolution using self-supervised deep learning
- Efficient Detection and Characterization of Targets of Natural Selection Using Transfer Learning
- MVAFormer: RGB-based Multi-View Spatio-Temporal Action Recognition with Transformer
- Condition Numbers and Eigenvalue Spectra of Shallow Networks on Spheres
- Neural Network Interoperability Across Platforms
- Object Detection as an Optional Basis: A Graph Matching Network for Cross-View UAV Localization
- Automatic Extraction of Road Networks by using Teacher-Student Adaptive Structural Deep Belief Network and Its Application to Landslide Disaster
- Energy-Efficient Deep Learning Without Backpropagation: A Rigorous Evaluation of Forward-Only Algorithms
- HyFormer-Net: A Synergistic CNN-Transformer with Interpretable Multi-Scale Fusion for Breast Lesion Segmentation and Classification in Ultrasound Images
- Parameter Interpolation Adversarial Training for Robust Image Classification
- Diluting Restricted Boltzmann Machines
- Beyond ImageNet: Understanding Cross-Dataset Robustness of Lightweight Vision Models
- Extended dynamic mode decomposition with dictionary learning: a data-driven adaptive spectral decomposition of the Koopman operator
- Image Hashing via Cross-View Code Alignment in the Age of Foundation Models
- Data-Augmented Deep Learning for Downhole Depth Sensing and Field Validation
- ODP-Bench: Benchmarking Out-of-Distribution Performance Prediction
- Exploring Landscapes for Better Minima along Valleys
- Integrating ConvNeXt and Vision Transformers for Enhancing Facial Age Estimation
- Spiking Neural Networks: The Future of Brain-Inspired Computing
- Comparative Analysis of Deep Learning Models for Olive Tree Crown and Shadow Segmentation Towards Biovolume Estimation
- EEG-Driven Image Reconstruction with Saliency-Guided Diffusion Models
- Generative Artificial Intelligence for Air Shower Simulation
- Learning Geometry: A Framework for Building Adaptive Manifold Models through Metric Optimization
- A Review of AI-Driven Approaches for Nanoscale Heat Conduction and Radiation
- DARTS: A Drone-Based AI-Powered Real-Time Traffic Incident Detection System
- Data-driven discovery of thermal illusions through latent-space geometry
- CHIPSIM: A Co-Simulation Framework for Deep Learning on Chiplet-Based Systems
- VISAT: Benchmarking Adversarial and Distribution Shift Robustness in Traffic Sign Recognition with Visual Attributes
- Feedback Alignment Meets Low-Rank Manifolds: A Structured Recipe for Local Learning
- PC-DARTS: Partial Channel Connections for Memory-Efficient Architecture Search
- Adversarial Domain Randomization
- Beyond Data Scarcity Optimizing R3GAN for Medical Image Generation from Small Datasets
- Deep Multi-Scale Features Learning for Distorted Image Quality Assessment
- Hybrid Models for Open Set Recognition
- Incremental Methods for Weakly Convex Optimization
- Deep Learning Algorithms with Applications to Video Analytics for A Smart City: A Survey
- Widget Captioning: Generating Natural Language Description for Mobile User Interface Elements
- Winner-Take-All Autoencoders
- Channel Pruning via Optimal Thresholding
- Hypergraph Convolution and Hypergraph Attention
- Deep Learning for Anomaly Detection: A Review
- Person Retrieval in Surveillance Video using Height, Color and Gender
- Medical Concept Representation Learning from Electronic Health Records and its Application on Heart Failure Prediction
- Sharpness-aware Quantization for Deep Neural Networks
- Federated Learning in Mobile Edge Networks: A Comprehensive Survey
- MetaFormer is Actually What You Need for Vision
- Understanding Information Processing in Human Brain by Interpreting Machine Learning Models
- Affinity and Diversity: Quantifying Mechanisms of Data Augmentation
- Deep Learning to Ternary Hash Codes by Continuation
- Camera-aware Proxies for Unsupervised Person Re-Identification
- Joint Object and State Recognition using Language Knowledge
- Twins: Revisiting the Design of Spatial Attention in Vision Transformers
- Deep causal representation learning for unsupervised domain adaptation
- Deep Network Classification by Scattering and Homotopy Dictionary Learning
- Uncertainty-Aware Attention for Reliable Interpretation and Prediction
- A Tandem Learning Rule for Effective Training and Rapid Inference of Deep Spiking Neural Networks
- CNN depth analysis with different channel inputs for Acoustic Scene Classification
- Amadeus: Scalable, Privacy-Preserving Live Video Analytics
- Pyramidal Convolution: Rethinking Convolutional Neural Networks for Visual Recognition
- Automatic Target Recognition on Synthetic Aperture Radar Imagery: A Survey
- Malleable 2.5D Convolution: Learning Receptive Fields along the Depth-axis for RGB-D Scene Parsing
- Growing a Brain: Fine-Tuning by Increasing Model Capacity
- Learning to Combine: Knowledge Aggregation for Multi-Source Domain Adaptation
- Edge AI: On-Demand Accelerating Deep Neural Network Inference via Edge Computing
- Adaptive Gradient for Adversarial Perturbations Generation
- Recent Deep Semi-supervised Learning Approaches and Related Works
- Bridging the Gap Between Spectral and Spatial Domains in Graph Neural Networks
- Deep Learning for the Classification of Lung Nodules
- Deep High-Resolution Representation Learning for Visual Recognition
- Discovering and Explaining the Representation Bottleneck of DNNs
- Face Completion with Semantic Knowledge and Collaborative Adversarial Learning
- MEAL: Multi-Model Ensemble via Adversarial Learning
- Regularizing Explanations in Bayesian Convolutional Neural Networks
- A Spatial-Temporal Decomposition Based Deep Neural Network for Time Series Forecasting
- Probabilistic Trust Intervals for Out of Distribution Detection
- Accelerating 3D Deep Learning with PyTorch3D
- Where am I looking at? Joint Location and Orientation Estimation by Cross-View Matching
- Neural Network Architectures for Location Estimation in the Internet of Things
- Multi Receptive Field Network for Semantic Segmentation
- Learning Neural Network Classifiers with Low Model Complexity
- Conditional Deep Learning for Energy-Efficient and Enhanced Pattern Recognition
- Tango: A Deep Neural Network Benchmark Suite for Various Accelerators
- Rethinking CNN Models for Audio Classification
- Character-level Chinese Writer Identification using Path Signature Feature, DropStroke and Deep CNN
- Voxceleb: Large-scale speaker verification in the wild
- Replay anti-spoofing countermeasure based on data augmentation with post selection
- Learning the Redundancy-free Features for Generalized Zero-Shot Object Recognition
- Gated Channel Transformation for Visual Recognition
- Variable Selection with Rigorous Uncertainty Quantification using Deep Bayesian Neural Networks: Posterior Concentration and Bernstein-von Mises Phenomenon
- Physical Attribute Prediction Using Deep Residual Neural Networks
- Weakly supervised object detection using pseudo-strong labels
- Universal Perturbation Attack Against Image Retrieval
- N-GCN: Multi-scale Graph Convolution for Semi-supervised Node Classification
- Temporal-Clustering Invariance in Irregular Healthcare Time Series
- RL-GAN-Net: A Reinforcement Learning Agent Controlled GAN Network for Real-Time Point Cloud Shape Completion
- On-Device Machine Learning: An Algorithms and Learning Theory Perspective
- Deep CTR Prediction in Display Advertising
- Dynamic Curriculum Learning for Imbalanced Data Classification
- Non-Negative Bregman Divergence Minimization for Deep Direct Density Ratio Estimation
- Harnessing Deep Neural Networks with Logic Rules
- Pyramid Vision Transformer: A Versatile Backbone for Dense Prediction without Convolutions
- Reclaiming AI as a theoretical tool for cognitive science
- FINCHES: A Computational Framework for Predicting Intermolecular Interactions in Intrinsically Disordered Proteins
- Hyperparameter Search in Machine Learning
- Fabric Surface Characterization: Assessment of Deep Learning-based Texture Representations Using a Challenging Dataset
- Representation Extraction and Deep Neural Recommendation for Collaborative Filtering
- On Hyperparameter Optimization of Machine Learning Algorithms: Theory and Practice
- NeST: A Neural Network Synthesis Tool Based on a Grow-and-Prune Paradigm
- A New Compensatory Genetic Algorithm-Based Method for Effective Compressed Multi-function Convolutional Neural Network Model Selection with Multi-Objective Optimization
- Inferring Algorithmic Patterns with Stack-Augmented Recurrent Nets
- Dynamic Graph CNN for Learning on Point Clouds
- Global Filter Networks for Image Classification
- ConvMLP: Hierarchical Convolutional MLPs for Vision
- Rethinking the Hyperparameters for Fine-tuning
- Chargrid-OCR: End-to-end Trainable Optical Character Recognition for Printed Documents using Instance Segmentation
- Places: An Image Database for Deep Scene Understanding
- Real World Robustness from Systematic Noise
- Object Detection in 20 Years: A Survey
- Cnns in land cover mapping with remote sensing imagery: a review and meta-analysis
- Open DNN Box by Power Side-Channel Attack
- A Statistician Teaches Deep Learning
- When Follow is Just One Click Away: Understanding Twitter Follow Behavior in the 2016 U.S. Presidential Election
- Adapting Auxiliary Losses Using Gradient Similarity
- A Deep Multi-task Learning Approach to Skin Lesion Classification
- Adaptive Selection of Deep Learning Models on Embedded Systems
- Dual-level Semantic Transfer Deep Hashing for Efficient Social Image Retrieval
- An interpretable deep hierarchical semantic convolutional neural network for lung nodule malignancy classification
- Continual Learning for Robotics: Definition, Framework, Learning Strategies, Opportunities and Challenges
- Spotting insects from satellites: modeling the presence of Culicoides imicola through Deep CNNs
- MeshingNet: A New Mesh Generation Method based on Deep Learning
- Ingredient-guided multi-modal interaction and refinement network for RGB-D food nutrition assessment
- Data augmentation using learned transformations for one-shot medical image segmentation
- Automatic Detection of Cerebral Microbleeds From MR Images via 3D Convolutional Neural Networks
- AggNet: Deep Learning From Crowds for Mitosis Detection in Breast Cancer Histology Images
- Fast Convolutional Neural Network Training Using Selective Data Sampling: Application to Hemorrhage Detection in Color Fundus Images
- Loss Landscape Dependent Self-Adjusting Learning Rates in Decentralized Stochastic Gradient Descent
- CoAtNet: Marrying Convolution and Attention for All Data Sizes
- For Manifold Learning, Deep Neural Networks can be Locality Sensitive Hash Functions
- Strategies for Pre-training Graph Neural Networks
- Optimal Feature Transport for Cross-View Image Geo-Localization
- An Efficient Multi-Scale Attention two-stream inflated 3D ConvNet network for cattle behavior recognition
- Shuffled Patch-Wise Supervision for Presentation Attack Detection
- Towards All-around Knowledge Transferring: Learning From Task-irrelevant Labels
- Unsupervised Learning of Solutions to Differential Equations with Generative Adversarial Networks
- Deep Long-Tailed Learning: A Survey
- Batch Group Normalization
- Split Slice Training Augmentation and Hyperparameter Tuning of RAKI Networks for Simultaneous Multi-Slice Reconstruction
- Revisiting Hierarchical Approach for Persistent Long-Term Video Prediction
- FlipReID: Closing the Gap between Training and Inference in Person Re-Identification
- Recent Advances in Neural Question Generation
- Efficient Semantic Scene Completion Network with Spatial Group Convolution
- MUREL: Multimodal Relational Reasoning for Visual Question Answering
- Facial Key Points Detection using Deep Convolutional Neural Network - NaimishNet
- Lifelong Graph Learning
- A Privacy-Preserving-Oriented DNN Pruning and Mobile Acceleration Framework
- EKT: Exercise-aware Knowledge Tracing for Student Performance Prediction
- Learning Spatiotemporal Features of Ride-sourcing Services with Fusion Convolutional Network
- When Residual Learning Meets Dense Aggregation: Rethinking the Aggregation of Deep Neural Networks
- Global Texture Enhancement for Fake Face Detection in the Wild
- A Simple Semi-Supervised Learning Framework for Object Detection
- FashionBERT: Text and Image Matching with Adaptive Loss for Cross-modal Retrieval
- Deep Residual Learning for Image Recognition
- BigGAN-based Bayesian reconstruction of natural images from human brain activity
- Deep 3D-to-2D Watermarking: Embedding Messages in 3D Meshes and Extracting Them from 2D Renderings
- HYPER-SNN: Towards Energy-efficient Quantized Deep Spiking Neural Networks for Hyperspectral Image Classification
- Fully Convolutional Networks for Multisource Building Extraction From an Open Aerial and Satellite Imagery Data Set
- Does Data Augmentation Benefit from Split BatchNorms
- Recognition of European mammals and birds in camera trap images using deep neural networks
- Do CNNs Encode Data Augmentations?
- Instance-Aware Predictive Navigation in Multi-Agent Environments
- How Transferable are CNN-based Features for Age and Gender Classification?
- DeepIris: Iris Recognition Using A Deep Learning Approach
- A system of different layers of abstraction for artificial intelligence
- Learning Optimal Data Augmentation Policies via Bayesian Optimization for Image Classification Tasks
- TorchIO: A Python library for efficient loading, preprocessing, augmentation and patch-based sampling of medical images in deep learning
- Sentiment Analysis from Images of Natural Disasters
- Distribution Alignment: A Unified Framework for Long-tail Visual Recognition
- Methods for Interpreting and Understanding Deep Neural Networks
- Robust Sparse Linear Discriminant Analysis
- Image Captioning and Visual Question Answering Based on Attributes and External Knowledge
- DiabDeep: Pervasive Diabetes Diagnosis based on Wearable Medical Sensors and Efficient Neural Networks
- The Rise of Radar for Autonomous Vehicles: Signal Processing Solutions and Future Research Directions
- Artificial Intelligence in the Rising Wave of Deep Learning: The Historical Path and Future Outlook [Perspectives]
- Algorithm Unrolling: Interpretable, Efficient Deep Learning for Signal and Image Processing
- Skeleton-Based Action Recognition With Gated Convolutional Neural Networks
- Perceptual Image Hashing for Content Authentication Based on Convolutional Neural Network With Multiple Constraints
- Omnidirectional Image Quality Assessment by Distortion Discrimination Assisted Multi-Stream Network
- Tensorizing Neural Networks
- GST: Group-Sparse Training for Accelerating Deep Reinforcement Learning
- Meta-Regularization by Enforcing Mutual-Exclusiveness
- Characterizing machine learning process: A maturity framework
- Circumventing Outliers of AutoAugment with Knowledge Distillation
- Information-Theoretic Understanding of Population Risk Improvement with Model Compression
- Bi-stream Pose Guided Region Ensemble Network for Fingertip Localization from Stereo Images
- Fold bifurcation identification through scientific machine learning
- Spectral Signatures in Backdoor Attacks
- A Dataset and Application for Facial Recognition of Individual Gorillas in Zoo Environments
- What makes visual place recognition easy or hard?
- GCNet: Non-local Networks Meet Squeeze-Excitation Networks and Beyond
- Deep Residual Dense U-Net for Resolution Enhancement in Accelerated MRI Acquisition
- DANCE: Enhancing saliency maps using decoys
- Competing Ratio Loss for Discriminative Multi-class Image Classification
- How Important is Weight Symmetry in Backpropagation?
- DenseCLIP: Language-Guided Dense Prediction with Context-Aware Prompting
- Learning Robust Representations via Multi-View Information Bottleneck
- Swin Transformer: Hierarchical Vision Transformer using Shifted Windows
- Dynamic Origin-Destination Matrix Prediction with Line Graph Neural Networks and Kalman Filter
- An Information Theory-inspired Strategy for Automatic Network Pruning
- Detecting COVID-19 from Breathing and Coughing Sounds using Deep Neural Networks
- Classification of Pathological and Normal Gait: A Survey
- Unifying Visual-Semantic Embeddings with Multimodal Neural Language Models
- Domain Adaptive SiamRPN++ for Object Tracking in the Wild
- Quantized Adam with Error Feedback
- A Theoretically Sound Upper Bound on the Triplet Loss for Improving the Efficiency of Deep Distance Metric Learning
- Machine Learning Etudes in Conformal Field Theories
- Noise-Sampling Cross Entropy Loss: Improving Disparity Regression Via Cost Volume Aware Regularizer
- Interpretable Deep Convolutional Fuzzy Classifier
- SWNet: Small-World Neural Networks and Rapid Convergence
- O sistema tecnológico digital
- LUTNet: Rethinking Inference in FPGA Soft Logic
- Cloud-based or On-device: An Empirical Study of Mobile Deep Inference
- Modulated binary cliquenet
- And the Bit Goes Down: Revisiting the Quantization of Neural Networks
- Training Deep Nets with Sublinear Memory Cost
- Uformer: A General U-Shaped Transformer for Image Restoration
- Controlling Recurrent Neural Networks by Conceptors
- Pseudo-ISP: Learning Pseudo In-camera Signal Processing Pipeline from A Color Image Denoiser
- Protein identification with deep learning: from abc to xyz
- Self-Supervised Multi-View Learning via Auto-Encoding 3D Transformations
- A Survey of Deep Learning Approaches for OCR and Document Understanding
- Measuring the Contribution of Multiple Model Representations in\n Detecting Adversarial Instances
- Dual-Sampling Attention Network for Diagnosis of COVID-19 from Community Acquired Pneumonia
- Efficient Reservoir Computing using Field Programmable Gate Array and Electro-optic Modulation
- Deep Reinforcement Learning with Quantum-inspired Experience Replay
- Multimodal Emotion Recognition for One-Minute-Gradual Emotion Challenge
- LiftPool: Bidirectional ConvNet Pooling
- SoftTriple Loss: Deep Metric Learning Without Triplet Sampling
- A Real-Time Cross-modality Correlation Filtering Method for Referring Expression Comprehension
- Exclusive: the most-cited papers of the twenty-first century
- Strategies for Conceptual Change in Convolutional Neural Networks
- PIXOR: Real-time 3D Object Detection from Point Clouds
- Hierarchical Neural Representation of Dreamed Objects Revealed by Brain Decoding with Deep Neural Network Features
- Testing Deep Learning Models for Image Analysis Using Object-Relevant Metamorphic Relations
- Pruning at a Glance: Global Neural Pruning for Model Compression
- Transform-Invariant Convolutional Neural Networks for Image Classification and Search
- Drop an Octave: Reducing Spatial Redundancy in Convolutional Neural Networks with Octave Convolution
- Contrastive Multiview Coding
- Exploiting Deep Features for Remote Sensing Image Retrieval: A Systematic Investigation
- CLCC: Contrastive Learning for Color Constancy
- Revisiting Adversarial Robustness Distillation: Robust Soft Labels Make Student Better
- Learning Channel Inter-dependencies at Multiple Scales on Dense Networks for Face Recognition
- VOLO: Vision Outlooker for Visual Recognition
- The Dynamicist Landscape
- CRIC: A VQA Dataset for Compositional Reasoning on Vision and Commonsense
- Multimodal mixing convolutional neural network and transformer for Alzheimer’s disease recognition
- Domain Generalization for Semantic Segmentation: A Survey
- Mitigating Spurious Correlation via Distributionally Robust Learning with Hierarchical Ambiguity Sets
- Self-Supervised Convolutional Subspace Clustering Network
- Stochastic Region Pooling: Make Attention More Expressive
- A Comprehensive Survey of Neural Architecture Search
- GPatt: Fast Multidimensional Pattern Extrapolation with Gaussian Processes
- CDGNet: Class Distribution Guided Network for Human Parsing
- Improving the Authentication with Built-in Camera Protocol Using Built-in Motion Sensors: A Deep Learning Solution
- Semi-supervised Domain Adaptation via Minimax Entropy
- Instance Adaptive Self-Training for Unsupervised Domain Adaptation
- Revisiting IM2GPS in the Deep Learning Era
- End-to-end Active Object Tracking and Its Real-world Deployment via Reinforcement Learning
- Adaptive Convolutional ELM For Concept Drift Handling in Online Stream Data
- A Survey on Bayesian Deep Learning
- Edge Intelligence: Paving the Last Mile of Artificial Intelligence with Edge Computing
- Calibration and Consistency of Adversarial Surrogate Losses
- CvS: Classification via Segmentation For Small Datasets
- A Sentiment-and-Semantics-Based Approach for Emotion Detection in Textual Conversations
- Towards Using Count-level Weak Supervision for Crowd Counting
- Unsupervised learning of clutter-resistant visual representations from natural videos
- Driver Drowsiness Detection Model Using Convolutional Neural Networks Techniques for Android Application
- Deep Frequent Spatial Temporal Learning for Face Anti-Spoofing
- Multi-Region Ensemble Convolutional Neural Network for Facial Expression Recognition
- Digital Twin: Values, Challenges and Enablers
- On the Importance of Visual Context for Data Augmentation in Scene\n Understanding
- Graph neural networks: A review of methods and applications
- Rotated Feature Network for multi-orientation object detection
- Defending against substitute model black box adversarial attacks with the 01 loss
- Demystifying Parallel and Distributed Deep Learning: An In-Depth Concurrency Analysis
- Noise Contrastive Priors for Functional Uncertainty
- Real-Time Steganalysis for Stream Media Based on Multi-channel Convolutional Sliding Windows
- WSOD2: Learning Bottom-up and Top-down Objectness Distillation for Weakly-supervised Object Detection
- A Battle of Network Structures: An Empirical Study of CNN, Transformer, and MLP
- RAP: Robustness-Aware Perturbations for Defending against Backdoor Attacks on NLP Models
- Video Coding for Machines: A Paradigm of Collaborative Compression and Intelligent Analytics
- Confusing Image Quality Assessment: Toward Better Augmented Reality Experience
- High Frequency Component Helps Explain the Generalization of Convolutional Neural Networks
- Generalizing from a Few Examples
- ROAM: Recurrently Optimizing Tracking Model
- The Multimodal Brain Tumor Image Segmentation Benchmark (BRATS)
- Acceleration of Deep Neural Network Training with Resistive Cross-Point Devices
- Deep Metric Learning for Practical Person Re-Identification
- Deep Lagrangian Networks: Using Physics as Model Prior for Deep Learning
- An Image Patch is a Wave: Phase-Aware Vision MLP
- Agriculture-Vision: A Large Aerial Image Database for Agricultural Pattern Analysis
- Any-Precision Deep Neural Networks
- Learning degraded image classification with restoration data fidelity
- Tomato plant disease classification using Multilevel Feature Fusion with adaptive channel spatial and pixel attention mechanism
- High-Performance Large-Scale Image Recognition Without Normalization
- A Semantics-Guided Class Imbalance Learning Model for Zero-Shot Classification
- Cross-modal Zero-shot Hashing
- Separating the Effects of Batch Normalization on CNN Training Speed and Stability Using Classical Adaptive Filter Theory
- SurveilEdge: Real-time Video Query based on Collaborative Cloud-Edge Deep Learning
- PAMS: Quantized Super-Resolution via Parameterized Max Scale
- Random VLAD based Deep Hashing for Efficient Image Retrieval
- RGB-based Semantic Segmentation Using Self-Supervised Depth Pre-Training
- Fine-Tuning Models Comparisons on Garbage Classification for Recyclability
- Improving Adversarial Transferability with Gradient Refining
- Generalized Out-of-Distribution Detection: A Survey
- A comprehensive review of object detection with deep learning
- Learning Various Length Dependence by Dual Recurrent Neural Networks
- Universal Lipschitz Approximation in Bounded Depth Neural Networks
- Image Synthesis with a Single (Robust) Classifier
- Symbol Emergence as an Interpersonal Multimodal Categorization
- Cooper: Cooperative Perception for Connected Autonomous Vehicles based on 3D Point Clouds
- In-domain representation learning for remote sensing
- Deep feature based rice leaf disease identification using support vector machine
- Sequential Graph Convolutional Network for Active Learning
- Performance Analysis and Characterization of Training Deep Learning Models on Mobile Devices
- Vector-quantized Image Modeling with Improved VQGAN
- Adversarial Domain Adaptation with Domain Mixup
- A Survey of Fake News: Fundamental Theories, Detection Methods, and Opportunities
- Video Modeling with Correlation Networks
- Reversed Active Learning based Atrous DenseNet for Pathological Image Classification
- Usefulness of interpretability methods to explain deep learning based plant stress phenotyping
- PathVQA: 30000+ Questions for Medical Visual Question Answering
- A Kronecker-factored approximate Fisher matrix for convolution layers
- Adaptive Future Frame Prediction with Ensemble Network
- Meta Feature Modulator for Long-tailed Recognition
- Stack-based Buffer Overflow Detection using Recurrent Neural Networks
- Beyond a Gaussian Denoiser: Residual Learning of Deep CNN for Image Denoising
- Pose-Invariant Embedding for Deep Person Re-Identification
- PlantDoc: A Dataset for Visual Plant Disease Detection
- Sim-to-Real Transfer of Robot Learning with Variable Length Inputs
- ActBERT: Learning Global-Local Video-Text Representations
- Deep learning in radiology: An overview of the concepts and a survey of the state of the art with focus on MRI
- End-to-End Blind Image Quality Assessment Using Deep Neural Networks
- Deep Learning Markov Random Field for Semantic Segmentation
- Post-Training Piecewise Linear Quantization for Deep Neural Networks
- Faster R-CNN: Towards Real-Time Object Detection with Region Proposal\n Networks
- DDLSTM: Dual-Domain LSTM for Cross-Dataset Action Recognition
- Learning with Group Invariant Features: A Kernel Perspective
- Intelligent Autofocus
- QuaterNet: A Quaternion-based Recurrent Model for Human Motion
- LightTrack: Finding Lightweight Neural Networks for Object Tracking via One-Shot Architecture Search
- A Simple Pooling-Based Design for Real-Time Salient Object Detection
- LightMOT: Lightweight and anchor-free solution for tracking multiple objects in dense populations
- On Learning Over-parameterized Neural Networks: A Functional Approximation Perspective
- Towards Spatial Variability Aware Deep Neural Networks (SVANN): A\n Summary of Results
- A Deep Neural Network for Audio Classification with a Classifier Attention Mechanism
- Deep High-Resolution Representation Learning for Visual Recognition
- PANNs: Large-Scale Pretrained Audio Neural Networks for Audio Pattern Recognition
- Improving Description-based Person Re-identification by Multi-granularity Image-text Alignments
- Learning to Localize: A 3D CNN Approach to User Positioning in Massive MIMO-OFDM Systems
- In-N-Out: Pre-Training and Self-Training using Auxiliary Information for Out-of-Distribution Robustness
- Material Recognition for Automated Progress Monitoring using Deep Learning Methods
- Data Extraction from Charts via Single Deep Neural Network
- Towards Rapid and Robust Adversarial Training with One-Step Attacks
- Low Precision Floating-point Arithmetic for High Performance FPGA-based CNN Acceleration
- Universal Person Re-Identification
- MAPS: A Synthetic Dataset for Probing Vision Models in a Controlled 3D Scene Space
- ResNet strikes back: An improved training procedure in timm
- Mandarin tone modeling using recurrent neural networks
- Deep Learning in Robotics: A Review of Recent Research
- Searching for A Robust Neural Architecture in Four GPU Hours
- Half and full solar cell efficiency binning by deep learning on electroluminescence images
- Deep Triplet Quantization
- Kernel-Based Smoothness Analysis of Residual Networks
- Toward Intelligent Sensing: Intermediate Deep Feature Compression
- Co-occurrence Feature Learning for Skeleton based Action Recognition using Regularized Deep LSTM Networks
- Stabilizing DARTS with Amended Gradient Estimation on Architectural Parameters
- Fantastic Four: Differentiable Bounds on Singular Values of Convolution Layers
- DeepACEv2: Automated Chromosome Enumeration in Metaphase Cell Images Using Deep Convolutional Neural Networks
- GenURL: A General Framework for Unsupervised Representation Learning
- Artistic Domain Generalisation Methods are Limited by their Deep Representations
- Feedback Attention for Cell Image Segmentation
- Dos and Don'ts of Machine Learning in Computer Security
- Attacking and Defending Machine Learning Applications of Public Cloud
- Natural Adversarial Examples
- Decision-based Universal Adversarial Attack
- Branchy-GNN: a Device-Edge Co-Inference Framework for Efficient Point Cloud Processing
- Spatio-temporal masked autoencoder-based phonetic segments classification from ultrasound
- Mix and Match: A Novel FPGA-Centric Deep Neural Network Quantization Framework
- MDCNN: A multimodal dual-CNN recursive model for fake news detection via audio- and text-based speech emotion recognition
- Optimizing Network Performance for Distributed DNN Training on GPU Clusters: ImageNet/AlexNet Training in 1.5 Minutes
- Can Single Neurons Solve MNIST? The Computational Power of Biological Dendritic Trees
- Improving Object Detection from Scratch via Gated Feature Reuse
- Successive Embedding and Classification Loss for Aerial Image Classification
- In-depth Question classification using Convolutional Neural Networks
- Feature Space Transfer for Data Augmentation
- Going Deeper for Multilingual Visual Sentiment Detection
- SAIA: Split Artificial Intelligence Architecture for Mobile Healthcare System
- Training Robust Deep Neural Networks via Adversarial Noise Propagation
- Salient Instance Segmentation via Subitizing and Clustering
- Deep Modulation Recognition with Multiple Receive Antennas: An End-to-end Feature Learning Approach
- Learning Multi-granular Quantized Embeddings for Large-Vocab Categorical Features in Recommender Systems
- Learning Pixel-level Semantic Affinity with Image-level Supervision for Weakly Supervised Semantic Segmentation
- WebFace260M: A Benchmark Unveiling the Power of Million-Scale Deep Face Recognition
- ESFNet: Efficient Network for Building Extraction from High-Resolution Aerial Images
- Self-Supervised Visual Feature Learning With Deep Neural Networks: A Survey
- A Review of Vibration-Based Damage Detection in Civil Structures: From Traditional Methods to Machine Learning and Deep Learning Applications
- DDcGAN: A Dual-Discriminator Conditional Generative Adversarial Network for Multi-Resolution Image Fusion
- Semantic-Aware Knowledge Preservation for Zero-Shot Sketch-Based Image Retrieval
- Convolutional Neural Network-based Topology Optimization (CNN-TO) By Estimating Sensitivity of Compliance from Material Distribution
- Rethinking Class-Balanced Methods for Long-Tailed Visual Recognition from a Domain Adaptation Perspective
- How to fine-tune deep neural networks in few-shot learning?
- Exploring the Limits of Language Modeling
- A Comprehensive Survey on Hardware-Aware Neural Architecture Search
- Learning Loss for Test-Time Augmentation
- Improving Network Slimming with Nonconvex Regularization
- Efficient Spatialtemporal Context Modeling for Action Recognition
- Universality of deep convolutional neural networks
- Improved Deep Convolutional Neural Network For Online Handwritten Chinese Character Recognition using Domain-Specific Knowledge
- BiSeNet V2: Bilateral Network with Guided Aggregation for Real-time Semantic Segmentation
- Siamese Natural Language Tracker: Tracking by Natural Language Descriptions with Siamese Trackers
- Multiresolution Convolutional Autoencoders
- Deep Convolutional Neural Network for Inverse Problems in Imaging
- SAWNet: A Spatially Aware Deep Neural Network for 3D Point Cloud Processing
- Neuroprosthesis for Decoding Speech in a Paralyzed Person with Anarthria
- Self-training with Noisy Student improves ImageNet classification
- Neural Networks Trained on Natural Scenes Exhibit Gestalt Closure
- Dank Learning: Generating Memes Using Deep Neural Networks
- Improving Prognostic Performance in Resectable Pancreatic Ductal Adenocarcinoma using Radiomics and Deep Learning Features Fusion in CT Images
- BlockDrop: Dynamic Inference Paths in Residual Networks
- When Person Re-identification Meets Changing Clothes
- Deep Facial Expression Recognition: A Survey
- StarGAN v2: Diverse Image Synthesis for Multiple Domains
- Deep Flow Collaborative Network for Online Visual Tracking
- An End-to-End Network for Panoptic Segmentation
- Continual World: A Robotic Benchmark For Continual Reinforcement Learning
- Combining Markov Random Fields and Convolutional Neural Networks for Image Synthesis
- Orthogonal Gradient Descent for Continual Learning
- Deep Adaptive Wavelet Network
- Bidirectional Mapping Generative Adversarial Networks for Brain MR to PET Synthesis
- A Survey of FPGA-Based Neural Network Accelerator
- Revisiting Self-Supervised Visual Representation Learning
- Deep Quaternion Features for Privacy Protection
- Unsupervised Learning of Invariant Representations in Hierarchical\n Architectures
- Room Geometry Estimation from Room Impulse Responses using Convolutional Neural Networks
- Learning to Hash for Indexing Big Data - A Survey
- CLDA: Contrastive Learning for Semi-Supervised Domain Adaptation
- Zero-shot World Models Are Developmentally Efficient Learners
- GhostNet: More Features From Cheap Operations
- Deep Inside Convolutional Networks: Visualising Image Classification Models and Saliency Maps
- DensePoint: Learning Densely Contextual Representation for Efficient Point Cloud Processing
- The Devil is in the Boundary: Exploiting Boundary Representation for Basis-based Instance Segmentation
- Reverse Transfer Learning: Can Word Embeddings Trained for Different NLP Tasks Improve Neural Language Models?
- Novel and Effective CNN-Based Binarization for Historically Degraded As-built Drawing Maps
- Multi-label Iterated Learning for Image Classification with Label Ambiguity
- ECA-Net: Efficient Channel Attention for Deep Convolutional Neural Networks
- HAMBox: Delving into Online High-quality Anchors Mining for Detecting Outer Faces
- Instance-weighted Central Similarity for Multi-label Image Retrieval
- Pairwise Teacher-Student Network for Semi-Supervised Hashing
- Real-world Mapping of Gaze Fixations Using Instance Segmentation for Road Construction Safety Applications
- Stanza: Layer Separation for Distributed Training in Deep Learning
- The OoO VLIW JIT Compiler for GPU Inference
- You Only Look & Listen Once: Towards Fast and Accurate Visual Grounding
- Convergence of a Relaxed Variable Splitting Coarse Gradient Descent Method for Learning Sparse Weight Binarized Activation Neural Networks
- Operational evaluation of data-driven forest fire forecasting models
- Escoin: Efficient Sparse Convolutional Neural Network Inference on GPUs
- Morphological Network: How Far Can We Go with Morphological Neurons?
- Spatial Pyramid Convolutional Neural Network for Social Event Detection in Static Image
- Data‐Efficient Electromagnetic Surrogate Solver Through Dissipative Relaxation Transfer Learning
- Intelligence Primer
- Salient Facial Features from Humans and Deep Neural Networks
- Enhancing Rotated Object Detection via Anisotropic Gaussian Bounding Box and Bhattacharyya Distance
- S2-MLP: Spatial-Shift MLP Architecture for Vision
- Distilling Image Classifiers in Object Detectors
- Content-Aware Convolutional Neural Networks
- Involution: Inverting the Inherence of Convolution for Visual Recognition
- AQD: Towards Accurate Fully-Quantized Object Detection
- Multi-Target Embodied Question Answering
- PUNCH: Positive UNlabelled Classification based information retrieval in Hyperspectral images
- Deep Virtual Networks for Memory Efficient Inference of Multiple Tasks
- Wavelet based edge feature enhancement for convolutional neural networks
- Improving Robustness Without Sacrificing Accuracy with Patch Gaussian Augmentation
- Classification of Motor Imagery EEG Signals by Using a Divergence Based Convolutional Neural Network
- Dropout as a Bayesian Approximation: Appendix
- Adaptive Normalized Risk-Averting Training For Deep Neural Networks
- Manipulating Identical Filter Redundancy for Efficient Pruning on Deep and Complicated CNN
- EIRES:Training-free AI-Generated Image Detection via Edit-Induced Reconstruction Error Shift
- DRIP: Dynamic patch Reduction via Interpretable Pooling
- Understanding Convolutional Neural Networks with Information Theory: An Initial Exploration
- Region Comparison Network for Interpretable Few-shot Image Classification
- Detection Defense Against Adversarial Attacks with Saliency Map
- FruitProm: Probabilistic Maturity Estimation and Detection of Fruits and Vegetables
- Corner Proposal Network for Anchor-free, Two-stage Object Detection
- Spatial prediction of apartment rent using regression-based and machine learning-based approaches with a large dataset
- TVT: Transferable Vision Transformer for Unsupervised Domain Adaptation
- Representation Learning: A Review and New Perspectives
- Two-Stream Convolutional Networks for Action Recognition in Videos
- Synergizing chemical and AI communities for advancing laboratories of the future
- Derivation and Analysis of Fast Bilinear Algorithms for Convolution
- Audio-visual Representation Learning for Anomaly Events Detection in Crowds
- Taming Visually Guided Sound Generation
- ResNet: Enabling Deep Convolutional Neural Networks through Residual Learning
- Seeking Salient Facial Regions for Cross-Database Micro-Expression Recognition
- Multi-view Low-rank Preserving Embedding: A Novel Method for Multi-view Representation
- Memory-based control with recurrent neural networks
- A representer theorem for deep neural networks
- LSUN: Construction of a Large-scale Image Dataset using Deep Learning with Humans in the Loop
- TAda! Temporally-Adaptive Convolutions for Video Understanding
- Ordinal Pooling Networks: For Preserving Information over Shrinking Feature Maps
- Digital video microscopy enhanced by deep learning
- SAND-mask: An Enhanced Gradient Masking Strategy for the Discovery of Invariances in Domain Generalization
- Benchmarking Neural Network Robustness to Common Corruptions and Surface Variations
- Deeply Learned Spectral Total Variation Decomposition
- Deep Learning for EEG motor imagery classification based on multi-layer CNNs feature fusion
- Deeply-learned and spatial–temporal feature engineering for human action understanding
- A backdoor attack against LSTM-based text classification systems
- Explaining Knowledge Distillation by Quantifying the Knowledge
- Image Categorization and Search via a GAT Autoencoder and Representative Models
- Hierarchically Robust Representation Learning
- Learned Dual-View Reflection Removal
- MRI-Based Brain Tumor Classification Using Ensemble of Deep Features and Machine Learning Classifiers
- Why Do Deep Residual Networks Generalize Better than Deep Feedforward Networks? -- A Neural Tangent Kernel Perspective
- PointGroup: Dual-Set Point Grouping for 3D Instance Segmentation
- ATP-Net: An Attention-based Ternary Projection Network For Compressed Sensing
- The Benchmarking Epistemology: Construct Validity for Evaluating Machine Learning Models
- Model-Behavior Alignment under Flexible Evaluation: When the Best-Fitting Model Isn't the Right One
- Privacy-Preserving Semantic Communication over Wiretap Channels with Learnable Differential Privacy
- Training a Convolutional Neural Network for Appearance-Invariant Place Recognition
- Common Task Framework For a Critical Evaluation of Scientific Machine Learning Algorithms
- Intrusion Detection: Machine Learning Baseline Calculations for Image Classification
- Rethinking Inference Placement for Deep Learning across Edge and Cloud Platforms: A Multi-Objective Optimization Perspective and Future Directions
- Self-Calibrated Consistency can Fight Back for Adversarial Robustness in Vision-Language Models
- SeeDNorm: Self-Rescaled Dynamic Normalization
- Deep Learning: A Critical Appraisal
- Multiple Convolutional Features in Siamese Networks for Object Tracking
- Multiclass Burn Wound Image Classification Using Deep Convolutional Neural Networks
- Quanvolutional Neural Networks for Pneumonia Detection: An Efficient Quantum-Assisted Feature Extraction Paradigm
- Quantum Machine Learning for Image Classification: A Hybrid Model of Residual Network with Quantum Support Vector Machine
- Looking for the Devil in the Details: Learning Trilinear Attention Sampling Network for Fine-grained Image Recognition
- Top-Down Semantic Refinement for Image Captioning
- Moving Beyond Diffusion: Hierarchy-to-Hierarchy Autoregression for fMRI-to-Image Reconstruction
- TrajGATFormer: A Graph-Based Transformer Approach for Worker and Obstacle Trajectory Prediction in Off-site Construction Environments
- An Improved Analysis of Stochastic Gradient Descent with Momentum
- Attention Residual Fusion Network with Contrast for Source-free Domain Adaptation
- Image Restoration Using Deep Regulated Convolutional Networks
- Energy-Efficient Domain-Specific Artificial Intelligence Models and Agents: Pathways and Paradigms
- From Black-box to Causal-box: Towards Building More Interpretable Models
- Toolflows for Mapping Convolutional Neural Networks on FPGAs: A Survey and Future Directions
- Task-Adaptive Neural Network Search with Meta-Contrastive Learning
- A Dilated Inception Network for Visual Saliency Prediction
- Widening and Squeezing: Towards Accurate and Efficient QNNs
- EBOP MAVEN: A machine learning model to estimate the input parameters for analytic fitting of detached eclipsing binary light curves
- OpenHype: Hyperbolic Embeddings for Hierarchical Open-Vocabulary Radiance Fields
- Bridging the gap to real-world language-grounded visual concept learning
- 3rd Place Solution to Large-scale Fine-grained Food Recognition
- 3rd Place Solution to ICCV LargeFineFoodAI Retrieval
- FINE Samples for Learning with Noisy Labels
- Parallelization Techniques for Verifying Neural Networks
- Long-tailed Species Recognition in the NACTI Wildlife Dataset
- Evaluating Bayesian Deep Learning Methods for Semantic Segmentation
- One-pixel Signature: Characterizing CNN Models for Backdoor Detection
- Stand-Alone Self-Attention in Vision Models
- Domain Generalization with MixStyle
- Communication-Efficient Federated Distillation
- xMem: A CPU-Based Approach for Accurate Estimation of GPU Memory in Deep Learning Training Workloads
- Graph Neural Regularizers for PDE Inverse Problems
- Memory Constrained Dynamic Subnetwork Update for Transfer Learning
- HybridSOMSpikeNet: A Deep Model with Differentiable Soft Self-Organizing Maps and Spiking Dynamics for Waste Classification
- Pose Augmentation: Class-agnostic Object Pose Transformation for Object Recognition
- A Semantic Loss Function for Deep Learning with Symbolic Knowledge
- Anderson-type acceleration method for Deep Neural Network optimization
- SAG-GAN: Semi-Supervised Attention-Guided GANs for Data Augmentation on Medical Images
- Machine learning identification of fractional-order vortex beam diffraction process
- Rethinking Cross-lingual Gaps from a Statistical Viewpoint
- Attentive Convolution: Unifying the Expressivity of Self-Attention with Convolutional Efficiency
- Stealing Neural Networks via Timing Side Channels
- Better Tokens for Better 3D: Advancing Vision-Language Modeling in 3D Medical Imaging
- TransTailor: Pruning the Pre-trained Model for Improved Transfer Learning
- Advances in Inference and Representation for Simultaneous Localization and Mapping
- ARC: A Vision-based Automatic Retail Checkout System
- Spatial Attention Point Network for Deep-learning-based Robust Autonomous Robot Motion Generation
- On the Power Saving in High-Speed Ethernet-based Networks for Supercomputers and Data Centers
- Study of Training Dynamics for Memory-Constrained Fine-Tuning
- BrainCognizer: Brain Decoding with Human Visual Cognition Simulation for fMRI-to-Image Reconstruction
- Precise classification of low quality G-banded Chromosome Images by reliability metrics and data pruning classifier
- Entity Embeddings of Categorical Variables
- A Semi-Parametric Estimation Method for the Quantile Spectrum with an Application to Earthquake Classification Using Convolutional Neural Network
- Layer-Parallel Training of Deep Residual Neural Networks
- Accelerated WGAN update strategy with loss change rate balancing
- A Unified Perspective on Optimization in Machine Learning and Neuroscience: From Gradient Descent to Neural Adaptation
- Lipschitz regularity of deep neural networks: analysis and efficient estimation
- Data-Driven Short-Term Voltage Stability Assessment Based on Spatial-Temporal Graph Convolutional Network
- Image augmentation with invertible networks in interactive satellite image change detection
- Visual Space Optimization for Zero-shot Learning
- Skin Cancer Recognition using Deep Residual Network
- Data-Driven Analysis of Intersectional Bias in Image Classification: A Framework with Bias-Weighted Augmentation
- Rethinking ResNets: Improved Stacking Strategies With High Order Schemes
- Privacy Inference Attacks and Defenses in Cloud-based Deep Neural Network: A Survey
- End-to-End Spoken Language Translation
- Towards In-Situ Failure Assessment: Deep Learning on DIC Results for Laminated Composites
- Large batch size training of neural networks with adversarial training and second-order information
- Learning from N-Tuple Data with M Positive Instances: Unbiased Risk Estimation and Theoretical Guarantees
- Deep Learning-Based Human Pose Estimation: A Survey
- MCANet: A Coherent Multimodal Collaborative Attention Network for Advanced Modulation Recognition in Adverse Noisy Environments
- BiSeNet: Bilateral Segmentation Network for Real-time Semantic Segmentation
- Application of Deep Neural Networks to assess corporate Credit Rating
- Decoding Dynamic Visual Experience from Calcium Imaging via Cell-Pattern-Aware Pretraining
- Provable Generalization Bounds for Deep Neural Networks with Momentum-Adaptive Gradient Dropout
- FreqPDE: Rethinking Positional Depth Embedding for Multi-View 3D Object Detection Transformers
- Deep Photovoltaic Nowcasting
- Modeling and Analysis of Energy Harvesting and Smart Grid-Powered Wireless Communication Networks: A Contemporary Survey
- Swin Transformer V2: Scaling Up Capacity and Resolution
- An Approximation of the Error Backpropagation Algorithm in a Predictive Coding Network with Local Hebbian Synaptic Plasticity
- Morphology-Aware KOA Classification: Integrating Graph Priors with Vision Models
- Automatic Classification of Circulating Blood Cell Clusters based on Multi-channel Flow Cytometry Imaging
- DRAMNet: Authentication based on Physical Unique Features of DRAM Using Deep Convolutional Neural Networks
- UMIche: A UMI-centric analysis platform for enhancing molecular quantification accuracy in bulk and single-cell sequencing
- Becoming episodic: The Development of Objectivity
- Application of CNN to a fine segmented scintillator detector for a single particle and neutrino-nucleon event
- Deep Lesion Tracker: Monitoring Lesions in 4D Longitudinal Imaging Studies
- Semantic-E2VID: a Semantic-Enriched Paradigm for Event-to-Video Reconstruction
- Symmetries in PAC-Bayesian Learning
- End-to-end Phoneme Sequence Recognition using Convolutional Neural Networks
- Quadratic Autoencoder (Q-AE) for Low-dose CT Denoising
- Themis: Fair and Efficient GPU Cluster Scheduling
- Fine-tuning Flow Matching Generative Models with Intermediate Feedback
- WP-CrackNet: A Collaborative Adversarial Learning Framework for End-to-End Weakly-Supervised Road Crack Detection
- Exploring Structural Degradation in Dense Representations for Self-supervised Learning
- Guiding Query Position and Performing Similar Attention for Transformer-Based Detection Heads
- View Adaptive Neural Networks for High Performance Skeleton-based Human Action Recognition
- ArmFormer: Lightweight Transformer Architecture for Real-Time Multi-Class Weapon Segmentation and Classification
- RSG: A Simple but Effective Module for Learning Imbalanced Datasets
- PyKale: Knowledge-Aware Machine Learning from Multiple Sources in Python
- FOX-NAS: Fast, On-device and Explainable Neural Architecture Search
- AUGUSTUS: An LLM-Driven Multimodal Agent System with Contextualized User Memory
- Enabling Data Diversity: Efficient Automatic Augmentation via Regularized Adversarial Training
- Moment-Based Domain Adaptation: Learning Bounds and Algorithms
- Rebooting Neuromorphic Hardware Design -- A Complexity Engineering Approach
- OpenEDS: Open Eye Dataset
- A deep learning approach to detecting volcano deformation from satellite imagery using synthetic datasets
- Video Swin Transformer
- CodeNet: Training Large Scale Neural Networks in Presence of Soft-Errors
- ELASTIC: Improving CNNs with Dynamic Scaling Policies
- Deep Variable-Block Chain with Adaptive Variable Selection
- Video-based Human Action Recognition using Deep Learning: A Review
- VM-BeautyNet: A Synergistic Ensemble of Vision Transformer and Mamba for Facial Beauty Prediction
- Deep generative priors for 3D brain analysis
- SimNets: A Generalization of Convolutional Networks
- A solution to generalized learning from small training sets found in infant repeated visual experiences of individual objects
- From Pixels to Words -- Towards Native Vision-Language Primitives at Scale
- Multi-modal video data-pipelines for machine learning with minimal human supervision
- Accelerating Minibatch Stochastic Gradient Descent using Typicality Sampling
- Tutorial on Variational Autoencoders
- Semantic representations emerge in biologically inspired ensembles of cross-supervising neural networks
- A Multi-domain Image Translative Diffusion StyleGAN for Iris Presentation Attack Detection
- Synergistic Integration and Discrepancy Resolution of Contextualized Knowledge for Personalized Recommendation
- Efficient Learning of Distributed Linear-Quadratic Controllers
- MUSE: Model-based Uncertainty-aware Similarity Estimation for zero-shot 2D Object Detection and Segmentation
- On the expressivity of sparse maxout networks
- Composition-based Multi-Relational Graph Convolutional Networks
- Scaling Vision Transformers for Functional MRI with Flat Maps
- DeDelayed: Deleting Remote Inference Delay via On-Device Correction
- Graph Neural Networks: A Review of Methods and Applications
- Deep learning-based prediction of response to HER2-targeted neoadjuvant chemotherapy from pre-treatment dynamic breast MRI: A multi-institutional validation study
- Accelerating Federated Learning via Momentum Gradient Descent
- O3BNN-R: An Out-of-Order Architecture for High-Performance and Regularized BNN Inference
- Multi-Agent Design Assistant for the Simulation of Inertial Fusion Energy
- Feature Denoising for Improving Adversarial Robustness
- The De-democratization of AI: Deep Learning and the Compute Divide in Artificial Intelligence Research
- Use of covariance matrix images for electroencephalography signal classification for multiclass motor imagery‐based brain computer interface
- Multi-Scale High-Resolution Logarithmic Grapher Module for Efficient Vision GNNs
- Randomness and Interpolation Improve Gradient Descent
- Plant leaf disease classification using EfficientNet deep learning model
- Stochastic Gradient/Mirror Descent: Minimax Optimality and Implicit Regularization
- Layer Normalization
- Visual7W: Grounded Question Answering in Images
- AnyUp: Universal Feature Upsampling
- GenCellAgent: Generalizable, Training-Free Cellular Image Segmentation via Large Language Model Agents
- Feature extraction of machine learning and phase transition point of Ising model
- Convolution, attention and structure embedding
- A Regularized Convolutional Neural Network for Semantic Image Segmentation
- Cautious Weight Decay
- MetaFormer Is Actually What You Need for Vision
- Metronome: Efficient Scheduling for Periodic Traffic Jobs with Network and Priority Awareness
- Local Background Features Matter in Out-of-Distribution Detection
- Revisiting Meta-Learning with Noisy Labels: Reweighting Dynamics and Theoretical Guarantees
- SDGraph: Multi-Level Sketch Representation Learning by Sparse-Dense Graph Architecture
- Faster Meta Update Strategy for Noise-Robust Deep Learning
- Tackling Catastrophic Forgetting and Background Shift in Continual Semantic Segmentation
- Robustness May Be at Odds with Accuracy
- Unsupervised Learning of Object Keypoints for Perception and Control
- Harnessing the Vulnerability of Latent Layers in Adversarially Trained Models
- R-Drop: Regularized Dropout for Neural Networks
- Lightweight CNN-Based Wi-Fi Intrusion Detection Using 2D Traffic Representations
- Cross Domain Image Matching in Presence of Outliers
- Fault Localization with Code Coverage Representation Learning
- Hierarchical Qubit-Merging Transformer for Quantum Error Correction
- High-resolution Photo Enhancement in Real-time: A Laplacian Pyramid Network
- Scalable Primitives for Generalized Sensor Fusion in Autonomous Vehicles
- Identification of tea foliar diseases and pest damage under practical field conditions using a convolutional neural network
- An Overview of Attacks and Defences on Intelligent Connected Vehicles
- Generating Binary Tags for Fast Medical Image Retrieval Based on Convolutional Nets and Radon Transform
- Adversarial Invariant Feature Learning with Accuracy Constraint for Domain Generalization
- Text-Enhanced Panoptic Symbol Spotting in CAD Drawings
- Source-Free Object Detection with Detection Transformer
- NeuroSwift: A Lightweight Cross-Subject Framework for fMRI Visual Reconstruction of Complex Scenes
- Weeping and Gnashing of Teeth: Teaching Deep Learning in Image and Video Processing Classes
- PERMDNN: Efficient Compressed DNN Architecture with Permuted Diagonal Matrices
- ΔEnergy: Optimizing Energy Change During Vision-Language Alignment Improves both OOD Detection and OOD Generalization
- Exact and Consistent Interpretation of Piecewise Linear Models Hidden behind APIs: A Closed Form Solution
- Statistical Guarantees for High-Dimensional Stochastic Gradient Descent
- A compressed code for memory discrimination
- Restricted Receptive Fields for Face Verification
- dN/dx Reconstruction with Deep Learning for High-Granularity TPCs
- High-Dimensional Learning Dynamics of Quantized Models with Straight-Through Estimator
- A Strategy of MR Brain Tissue Images' Suggestive Annotation Based on Modified U-Net
- Deep learning methods based on cross-section images for predicting\n effective thermal conductivity of composites
- Learning the mapping x↦ ∑i=1d xi2: the cost of finding the needle in a haystack
- GroupFormer: Group Activity Recognition with Clustered Spatial-Temporal Transformer
- Can a powerful neural network be a teacher for a weaker neural network?
- CryptGPU: Fast Privacy-Preserving Machine Learning on the GPU
- Distributed Learning of Deep Neural Networks using Independent Subnet Training
- Automated machine learning: Review of the state-of-the-art and opportunities for healthcare
- Sketch Animation: State-of-the-art Report
- Understanding Notions of Stationarity in Non-Smooth Optimization
- A Large-Scale Benchmark for Food Image Segmentation
- Deep Learning using Linear Support Vector Machines
- Translution: Unifying Self-attention and Convolution for Adaptive and Relative Modeling
- Local-Global Context-Aware and Structure-Preserving Image Super-Resolution
- Automated Glaucoma Report Generation via Dual-Attention Semantic Parallel-LSTM and Multimodal Clinical Data Integration
- Stochastic Training is Not Necessary for Generalization
- Prismo: A Decision Support System for Privacy-Preserving ML Framework Selection
- Probabilistic Variational Contrastive Learning
- Machine learning in ant biology research: A systematic review
- Center-Focusing Multi-task CNN with Injected Features for Classification of Glioma Nuclear Images
- Small is Sufficient: Reducing the World AI Energy Consumption Through Model Selection
- Artificial intelligence for education: Knowledge and its assessment in AI-enabled learning ecologies
- Computer-aided diagnosis of endobronchial ultrasound images using convolutional neural network
- Self-taught Object Localization with Deep Networks
- Understanding Character Recognition using Visual Explanations Derived from the Human Visual System and Deep Networks
- SimpleDet: A Simple and Versatile Distributed Framework for Object Detection and Instance Recognition
- DeepHash: Getting Regularization, Depth and Fine-Tuning Right
- Piecewise Linear Units Improve Deep Neural Networks
- OR-Net: Pointwise Relational Inference for Data Completion under Partial Observation
- CAMAL: Context-Aware Multi-layer Attention framework for Lightweight Environment Invariant Visual Place Recognition
- Deep Neural Networks for Choice Analysis: Architectural Design with Alternative-Specific Utility Functions
- Architecture Induces Structural Invariant Manifolds of Neural Network Training Dynamics
- Deep prior-based denoising for state-of-the-art scientific imaging and metrology
- Needles in Haystacks: On Classifying Tiny Objects in Large Images
- Drill the Cork of Information Bottleneck by Inputting the Most Important Data
- SpineNet: Learning Scale-Permuted Backbone for Recognition and Localization
- Fully Convolutional Networks for Semantic Segmentation
- AS-MLP: An Axial Shifted MLP Architecture for Vision
- Learning Transferable Adversarial Examples via Ghost Networks
- AGDC: Automatic Garbage Detection and Collection
- Provable Watermarking for Data Poisoning Attacks
- Distributionally robust approximation property of neural networks
- Robustness of Object Recognition under Extreme Occlusion in Humans and Computational Models
- Integrating Specialized Classifiers Based on Continuous Time Markov Chain
- MAT-Agent: Adaptive Multi-Agent Training Optimization
- MNIST-NET10: A heterogeneous deep networks fusion based on the degree of certainty to reach 0.1 error rate. Ensembles overview and proposal
- VirDA: Reusing Backbone for Unsupervised Domain Adaptation with Visual Reprogramming
- Self-Training With Noisy Student Improves ImageNet Classification
- RAVEN: A Dataset for Relational and Analogical Visual rEasoNing
- LieTransformer: Equivariant self-attention for Lie Groups
- Exploring Data Aggregation and Transformations to Generalize across Visual Domains
- Image Classification with Classic and Deep Learning Techniques
- The impact of abstract and object tags on image privacy classification
- DISCO: Diversifying Sample Condensation for Efficient Model Evaluation
- Deep Neural Networks Inspired by Differential Equations
- C-Net: A Reliable Convolutional Neural Network for Biomedical Image Classification
- Exploring Modality-shared Appearance Features and Modality-invariant Relation Features for Cross-modality Person Re-Identification
- DeepSpectrumLite: A Power-Efficient Transfer Learning Framework for Embedded Speech and Audio Processing from Decentralised Data
- Rectifier Neural Network with a Dual-Pathway Architecture for Image Denoising
- Delta-STN: Efficient Bilevel Optimization for Neural Networks using Structured Response Jacobians
- Joint Architecture and Knowledge Distillation in CNN for Chinese Text Recognition
- Sparse components distinguish visual pathways & their alignment to neural networks
- TinySpeech: Attention Condensers for Deep Speech Recognition Neural Networks on Edge Devices
- What Do You See? Evaluation of Explainable Artificial Intelligence (XAI) Interpretability through Neural Backdoors
- SGAS: Sequential Greedy Architecture Search
- Representation Learning from Limited Educational Data with Crowdsourced Labels
- Evaluation of Model Selection for Kernel Fragment Recognition in Corn Silage
- Anomaly Detection in Univariate Time-series: A Survey on the State-of-the-Art
- Time-Frequency Analysis based Blind Modulation Classification for Multiple-Antenna Systems
- AIM 2020 Challenge on Video Extreme Super-Resolution: Methods and Results
- Learning Longterm Representations for Person Re-Identification Using Radio Signals
- A Unified Object Motion and Affinity Model for Online Multi-Object Tracking
- Temporal Accumulative Features for Sign Language Recognition
- Long-Tailed Recognition via Information-Preservable Two-Stage Learning
- Quick-CapsNet (QCN): A fast alternative to Capsule Networks
- Artificial Hippocampus Networks for Efficient Long-Context Modeling
- Label-frugal satellite image change detection with generative virtual exemplar learning
- Interactive reconstruction of Monte Carlo image sequences using a recurrent denoising autoencoder
- The dilemma of quantum neural networks
- Dataset Reuse: Toward Translating Principles to Practice
- A review of evolving remote sensing and automated techniques in rock glacier mapping
- ImageNet Large Scale Visual Recognition Challenge
- Random Erasing Data Augmentation
- CoEdge: Cooperative DNN Inference with Adaptive Workload Partitioning over Heterogeneous Edge Devices
- Rapid computation of high-level visual surprise
- SympNets: Intrinsic structure-preserving symplectic networks for identifying Hamiltonian systems
- Verifying Memoryless Sequential Decision-making of Large Language Models
- Associative Memory Model with Neural Networks: Memorizing multiple images with one neuron
- Shaken or Stirred? An Analysis of MetaFormer's Token Mixing for Medical Imaging
- TreeNet: Layered Decision Ensembles
- Riddled basin geometry sets fundamental limits to predictability and reproducibility in deep learning
- Multi-task Neural Networks for QSAR Predictions
- Learning the detector in optical tomography
- Computing frustration and near-monotonicity in deep neural networks
- Evaluation of fish feeding intensity in aquaculture using a convolutional neural network and machine vision
- Neuroplastic Modular Framework: Cross-Domain Image Classification of Garbage and Industrial Surfaces
- Beyond Random: Automatic Inner-loop Optimization in Dataset Distillation
- Visual Representations inside the Language Model
- Bridging Reasoning to Learning: Unmasking Illusions using Complexity Out of Distribution Generalization
- Improving Large-Scale Recommender Systems with Auxiliary Learning
- Exploring the Efficacy of Modified Transfer Learning in Identifying Parkinson's Disease Through Drawn Image Patterns
- POMO: Policy Optimization with Multiple Optima for Reinforcement Learning
- Improving Transferability of Adversarial Examples with Input Diversity
- Kaleidoscope: An Efficient, Learnable Representation For All Structured Linear Maps
- Continuous vs. Discrete Optimization of Deep Neural Networks
- FastFCN: Rethinking Dilated Convolution in the Backbone for Semantic Segmentation
- Multi-Scale Aligned Distillation for Low-Resolution Detection
- A Mathematical Explanation of Transformers for Large Language Models and GPTs
- Using predefined vector systems as latent space configuration for neural network supervised training on data with arbitrarily large number of classes
- Attention Branch Network: Learning of Attention Mechanism for Visual Explanation
- Quantization Range Estimation for Convolutional Neural Networks
- Interpolated Convolutional Networks for 3D Point Cloud Understanding
- ConvBERT: Improving BERT with Span-based Dynamic Convolution
- Evaluation of Transfer Learning for Classification of: (1) Diabetic Retinopathy by Digital Fundus Photography and (2) Diabetic Macular Edema, Choroidal Neovascularization and Drusen by Optical Coherence Tomography
- End-to-end Active Object Tracking via Reinforcement Learning
- Generation and Comprehension of Unambiguous Object Descriptions
- Keeping Your Eye on the Ball: Trajectory Attention in Video Transformers
- Task-Level Contrastiveness for Cross-Domain Few-Shot Learning
- HAVIR: HierArchical Vision to Image Reconstruction using CLIP-Guided Versatile Diffusion
- Conditional Pseudo-Supervised Contrast for Data-Free Knowledge Distillation
- Optimal Rates for Generalization of Gradient Descent for Deep ReLU Classification
- Exploring Accuracy Law for Deep Time Series Forecasters: An Empirical Study
- Hyperparameter Loss Surfaces Are Simple Near their Optima
- Using Fourier Analysis and Mutant Clustering to Accelerate DNN Mutation Testing
- Visual Language Model as a Judge for Object Detection in Industrial Diagrams
- An Efficient Quality Metric for Video Frame Interpolation Based on Motion-Field Divergence
- Image Generation Based on Image Style Extraction
- Real-time Hand Gesture Detection and Classification Using Convolutional Neural Networks
- Fast Object Detection in Compressed Video
- Phase Collaborative Network for Two-Phase Medical Image Segmentation
- TextCAM: Explaining Class Activation Map with Text
- GLAI: GreenLightningAI for Accelerated Training through Knowledge Decoupling
- Latency-Aware Differentiable Neural Architecture Search
- Realization of spatial sparseness by deep ReLU nets with massive data
- Indices Matter: Learning to Index for Deep Image Matting
- Semantic Visual Simultaneous Localization and Mapping: A Survey on State of the Art, Challenges, and Future Directions
- AWAC: Accelerating Online Reinforcement Learning with Offline Datasets
- Max-Pooling Dropout for Regularization of Convolutional Neural Networks
- Hessian-Aware Pruning and Optimal Neural Implant
- Robust Context-Aware Object Recognition
- Signal Classification Recovery Across Domains Using Unsupervised Domain Adaptation
- Assessing Foundation Models for Mold Colony Detection with Limited Training Data
- Quantum Probabilistic Label Refining: Enhancing Label Quality for Robust Image Classification
- FourPhononGPU: A GPU-accelerated framework for calculating phonon scattering rates and thermal conductivity
- Uformer: A General U-Shaped Transformer for Image Restoration
- Digital Passport: A Novel Technological Strategy for Intellectual Property Protection of Convolutional Neural Networks
- Normal-Abnormal Guided Generalist Anomaly Detection
- On-the-Fly Data Augmentation via Gradient-Guided and Sample-Aware Influence Estimation
- Advances in Medical Image Segmentation: A Comprehensive Survey with a Focus on Lumbar Spine Applications
- Finding beans in burgers: Deep semantic-visual embedding with localization
- PointConv: Deep Convolutional Networks on 3D Point Clouds
- DynamicViT: Efficient Vision Transformers with Dynamic Token Sparsification
- PhraseStereo: The First Open-Vocabulary Stereo Image Segmentation Dataset
- On zero-shot recognition of generic objects
- Time Series Forecasting With Deep Learning: A Survey
- Deep Neural Networks are Easily Fooled: High Confidence Predictions for\n Unrecognizable Images
- From Videos to Indexed Knowledge Graphs -- Framework to Marry Methods for Multimodal Content Analysis and Understanding
- Beyond the Memory Wall: A Case for Memory-centric HPC System for Deep Learning
- FARSA: Fully Automated Roadway Safety Assessment
- Continuously Augmented Discrete Diffusion model for Categorical Generative Modeling
- AIRCHITECT: Learning Custom Architecture Design and Mapping Space
- Benchmarking Deep Learning Convolutions on Energy-constrained CPUs
- From MNIST to ImageNet: Understanding the Scalability Boundaries of Differentiable Logic Gate Networks
- The Impact of Scaling Training Data on Adversarial Robustness
- Sharpness of Minima in Deep Matrix Factorization: Exact Expressions
- Using Images from a Video Game to Improve the Detection of Truck Axles
- Effective Model Pruning
- Hybrid Dual-Batch and Cyclic Progressive Learning for Efficient Distributed Training
- AttentionViG: Cross-Attention-Based Dynamic Neighbor Aggregation in Vision GNNs
- Human vs. AI Safety Perception? Decoding Human Safety Perception with Eye-Tracking Systems, Street View Images, and Explainable AI
- Spontaneous High-Order Generalization in Neural Theory-of-Mind Networks
- Accelerating Dynamic Image Graph Construction on FPGA for Vision GNNs
- Symmetry-Aware Bayesian Optimization via Max Kernels
- CLASP: Adaptive Spectral Clustering for Unsupervised Per-Image Segmentation
- Automatic recognition of feeding and foraging behaviour in pigs using deep learning
- ThermalGen: Style-Disentangled Flow-Based Generative Models for RGB-to-Thermal Image Translation
- Adaptive Precision Training: Quantify Back Propagation in Neural Networks with Fixed-point Numbers
- Enabling Physical AI through Biological Principles
- Specialization after Generalization: Towards Understanding Test-Time Training in Foundation Models
- DRIFT: Divergent Response in Filtered Transformations for Robust Adversarial Defense
- PEARL: Performance-Enhanced Aggregated Representation Learning
- Towards Foundation Models for Cryo-ET Subtomogram Analysis
- High-Order Progressive Trajectory Matching for Medical Image Dataset Distillation
- BALF: Budgeted Activation-Aware Low-Rank Factorization for Fine-Tuning-Free Model Compression
- What Are We Automating? On the Need for Vision and Expertise When Deploying AI Systems
- Plant3R: Fusing 3D feature learning with Gaussian splatting to enhance wheat plant 3D reconstruction precision
- MLPerf Training Benchmark
- Learning to Diversify for Single Domain Generalization
- The Devil is in the Margin: Margin-based Label Smoothing for Network Calibration
- Model Watermarking for Image Processing Networks
- A Consolidated Approach to Convolutional Neural Networks and the Kolmogorov Complexity
- Retrieve-Then-Adapt: Example-based Automatic Generation for Proportion-related Infographics
- Road Crack Detection Using Deep Convolutional Neural Network and Adaptive Thresholding
- Exploring Uncertainty in Deep Learning for Construction of Prediction Intervals
- Semantically-Aware Strategies for Stereo-Visual Robotic Obstacle Avoidance
- Tent: Fully Test-time Adaptation by Entropy Minimization
- LifeCLEF Plant Identification Task 2014
- LifeCLEF Plant Identification Task 2014
- Gradient Flow Convergence Guarantee for General Neural Network Architectures
- FairViT-GAN: A Hybrid Vision Transformer with Adversarial Debiasing for Fair and Explainable Facial Beauty Prediction
- Influence-Guided Concolic Testing of Transformer Robustness
- Towards Interpretable Visual Decoding with Attention to Brain Representations
- A Computational Perspective on NeuroAI and Synthetic Biological Intelligence
- Spatially Parallel All-optical Neural Networks
- QuadEnhancer: Leveraging Quadratic Transformations to Enhance Deep Neural Networks
- Model-agnostic interpretation by visualization of feature perturbations
- Probabilistic Bearing Fault Diagnosis Using Gaussian Process with Tailored Feature Extraction
- Deep Learning Approaches with Explainable AI for Differentiating Alzheimer Disease and Mild Cognitive Impairment
- The unbearable slowness of being: Why do we live at 10 bits/s?
- Machine learning for synthetic gene circuit engineering
- Foundation models in bioinformatics
- Patch Rebirth: Toward Fast and Transferable Model Inversion of Vision Transformers
- Sparse, Collaborative, or Nonnegative Representation: Which Helps Pattern Classification?
- Leader Stochastic Gradient Descent for Distributed Training of Deep Learning Models: Extension
- Program-Guided Image Manipulators
- Metric-based Regularization and Temporal Ensemble for Multi-task Learning using Heterogeneous Unsupervised Tasks
- Stochastic Interpolants via Conditional Dependent Coupling
- HTMA-Net: Towards Multiplication-Avoiding Neural Networks via Hadamard Transform and In-Memory Computing
- DPFNAS: Differential Privacy-Enhanced Federated Neural Architecture Search for 6G Edge Intelligence
- Learning Modulated Loss for Rotated Object Detection
- Learning Generalizable Visual Representations via Interactive Gameplay
- CProp: Adaptive Learning Rate Scaling from Past Gradient Conformity
- On Controlled DeEntanglement for Natural Language Processing
- Deep Learning for Oral Health: Benchmarking ViT, DeiT, BEiT, ConvNeXt, and Swin Transformer
- Adaptive and Iteratively Improving Recurrent Lateral Connections
- How Secure is Distributed Convolutional Neural Network on IoT Edge Devices?
- Targeted perturbations reveal brain-like local coding axes in robustified, but not standard, ANN-based brain models
- A Unified Approximation Framework for Compressing and Accelerating Deep Neural Networks
- MindCraft: How Concept Trees Take Shape In Deep Models
- PAPER: Privacy-Preserving Convolutional Neural Networks using Low-Degree Polynomial Approximations and Structural Optimizations on Leveled FHE
- Survey of Machine Learning Accelerators
- IONext: Unlocking the Next Era of Inertial Odometry
- A Sparse CNN Accelerator for Eliminating Redundant Computations in Intra- and Inter-Convolutional/Pooling Layers
- TSDM: Tracking by SiamRPN++ with a Depth-refiner and a Mask-generator
- Comparison and Benchmarking of AI Models and Frameworks on Mobile Devices
- Generative Image Modeling Using Spatial LSTMs
- Understand Scene Categories by Objects: A Semantic Regularized Scene Classifier Using Convolutional Neural Networks
- Visual Relationship Detection using Scene Graphs: A Survey
- Restructuring Batch Normalization to Accelerate CNN Training
- Mining Domain Knowledge: Improved Framework towards Automatically Standardizing Anatomical Structure Nomenclature in Radiotherapy
- A Learning-from-noise Dilated Wide Activation Network for denoising Arterial Spin Labeling (ASL) Perfusion Images
- On Embeddings in Relational Databases
- deepSELF: An Open Source Deep Self End-to-End Learning Framework
- Reconstructed spatial receptive field structures by reverse correlation technique explains the visual feature selectivity of units in deep convolutional neural networks
- SocialAI 0.1: Towards a Benchmark to Stimulate Research on Socio-Cognitive Abilities in Deep Reinforcement Learning Agents
- Design of Reconfigurable Multi-Operand Adder for Massively Parallel Processing
- Spectral Collapse Drives Loss of Plasticity in Deep Continual Learning
- Progressive Weight Loading: Accelerating Initial Inference and Gradually Boosting Performance on Resource-Constrained Environments
- Task-Adaptive Parameter-Efficient Fine-Tuning for Weather Foundation Models
- Light Differentiable Logic Gate Networks
- Prophecy: Inferring Formal Properties from Neuron Activations
- A Data-driven Typology of Vision Models from Integrated Representational Metrics
- LANCE: Low Rank Activation Compression for Efficient On-Device Continual Learning
- Temporal vs. Spatial: Comparing DINOv3 and V-JEPA2 Feature Representations for Video Action Analysis
- AI for Sustainable Future Foods
- Mammo-CLIP Dissect: A Framework for Analysing Mammography Concepts in Vision-Language Models
- Deep Adversarially-Enhanced k-Nearest Neighbors
- Short-term Load Forecasting with Deep Residual Networks
- Plant identification based on noisy web data: the amazing performance of deep learning (LifeCLEF 2017)
- DeepEMD: Differentiable Earth Mover's Distance for Few-Shot Learning
- Grad-CAM: Visual Explanations from Deep Networks via Gradient-based Localization
- Learning Phrase Representations using RNN Encoder-Decoder for Statistical Machine Translation
- The Taboo Trap: Behavioural Detection of Adversarial Samples
- Explanations can be manipulated and geometry is to blame
- MeshCNN: A Network with an Edge
- ACNet: Strengthening the Kernel Skeletons for Powerful CNN via Asymmetric Convolution Blocks
- Data Augmentation for Skin Lesion using Self-Attention based Progressive Generative Adversarial Network
- Audio-Visual Transformer Based Crowd Counting
- Preparation Meets Opportunity: Enhancing Data Preprocessing for ML Training With Seneca
- On the basis of brain: neural-network-inspired changes in general-purpose chips
- Supervised Contrastive Learning
- Unleashing the Potential of the Semantic Latent Space in Diffusion Models for Image Dehazing
- Embodied AI: From LLMs to World Models
- GridMask Data Augmentation
- Impact of Loss Weight and Model Complexity on Physics-Informed Neural Networks for Computational Fluid Dynamics
- Sobolev acceleration for neural networks
- Mamba Modulation: On the Length Generalization of Mamba
- Adaptive von Mises-Fisher Likelihood Loss for Supervised Deep Time Series Hashing
- PAC-Bayes Analysis of Sentence Representation
- Learning to see across Domains and Modalities
- Thickened 2D Networks for Efficient 3D Medical Image Segmentation
- Not Only Look But Observe: Variational Observation Model of Scene-Level 3D Multi-Object Understanding for Probabilistic SLAM
- Deep Control - a simple automatic gain control for memory efficient and high performance training of deep convolutional neural networks
- Energy-Efficient Processing and Robust Wireless Cooperative Transmission for Edge Inference
- AutoGrow: Automatic Layer Growing in Deep Convolutional Networks
- Seesaw-Net: Convolution Neural Network With Uneven Group Convolution
- Structure-Aware Face Clustering on a Large-Scale Graph with \bf107 Nodes
- Unsupervised Domain Adaptation for Object Detection via Cross-Domain Semi-Supervised Learning
- Multi-task Learning by Leveraging the Semantic Information
- Understanding Submodular Information Measure Based Objectives for Representation Learning: A Variance and Separation Perspective
- Deep(er) Learning
- MSG-Transformer: Exchanging Local Spatial Information by Manipulating Messenger Tokens
- Hybrid Composition with IdleBlock: More Efficient Networks for Image Recognition
- Temporal Poisoning: Clean-Label Backdoors via Event Redistribution in SNNs
- Deep learning for brake squeal: vibration detection, characterization and prediction
- Multi Layer Neural Networks as Replacement for Pooling Operations
- Lets keep it simple, Using simple architectures to outperform deeper and\n more complex architectures
- Explorative Modeling: Unlocking a Third Pretraining Axis and End-to-End Generation
- Large-Scale Long-Tailed Recognition in an Open World
- Thinking in Scales: Accelerating Gigapixel Pathology Image Analysis via Adaptive Continuous Reasoning
- Artificial Intelligence Empowered New Materials: Discovery, Synthesis, Prediction to Validation
- Does the brain's ventral visual pathway compute object shape?
- Application of MobileNet and Xception neural networks to identify Sillago sihama populations in Vietnam's coastal waters based on otolith morphology
- Deep learning in plant phenotyping: the first ten years
- Recurrence along Depth: Deep Convolutional Neural Networks with Recurrent Layer Aggregation
- Quantitative Evaluations on Saliency Methods: An Experimental Study
- Deep learning‐based association analysis of root image data and cucumber yield
- Unifying Relational Sentence Generation and Retrieval for Medical Image Report Composition
- Feature refinement: An expression-specific feature learning and fusion method for micro-expression recognition
- Ellipse Regression with Predicted Uncertainties for Accurate Multi-View 3D Object Estimation
- Multifaceted Feature Visualization: Uncovering the Different Types of Features Learned By Each Neuron in Deep Neural Networks
- DAIL: Dataset-Aware and Invariant Learning for Face Recognition
- When AI Meets Science: Research Diversity, Interdisciplinarity, Visibility, and Retractions across Disciplines in a Global Surge
- T-BFA: Targeted Bit-Flip Adversarial Weight Attack
- Multi-path Neural Networks for On-device Multi-domain Visual Classification
- Channel Tiling for Improved Performance and Accuracy of Optical Neural Network Accelerators
- Dynamic DNN Decomposition for Lossless Synergistic Inference
- Accuracy and Architecture Studies of Residual Neural Network solving Ordinary Differential Equations
- Learning Curves for Analysis of Deep Networks
- Dream-Cubed: Controllable Generative Modeling in Minecraft by Training on Billions of Cubes
- Semantic Understanding of Scenes Through the ADE20K Dataset
- PotentialNet for Molecular Property Prediction
- CLIP-Adapter: Better Vision-Language Models with Feature Adapters
- Estimation error analysis of deep learning on the regression problem on the variable exponent Besov space
- Selecting Data Augmentation for Simulating Interventions
- Gesture Recognition for Initiating Human-to-Robot Handovers
- Composition-Aware Image Aesthetics Assessment
- Unsupervised Intuitive Physics from Past Experiences
- Fuzzy Semantic Segmentation of Breast Ultrasound Image with Breast Anatomy Constraints
- Medical Image Segmentation Using a U-Net type of Architecture
- Thanks for Nothing: Predicting Zero-Valued Activations with Lightweight Convolutional Neural Networks
- A Deep Learning Framework for Classification of in vitro Multi-Electrode Array Recordings
- Inductive Bias of Gradient Descent based Adversarial Training on Separable Data
- Zero-shifting Technique for Deep Neural Network Training on Resistive Cross-point Arrays
- Compact deep neural network models of the visual cortex
- Vision Permutator: A Permutable MLP-Like Architecture for Visual Recognition
- SEALing Neural Network Models in Secure Deep Learning Accelerators
- Quantum Energy Regression using Scattering Transforms
- Exact and Consistent Interpretation for Piecewise Linear Neural Networks: A Closed Form Solution
- Dense Adaptive Cascade Forest: A Self Adaptive Deep Ensemble for Classification Problems
- Gotta Adapt 'Em All: Joint Pixel and Feature-Level Domain Adaptation for Recognition in the Wild
- k-Nearest Neighbors by Means of Sequence to Sequence Deep Neural Networks and Memory Networks
- Application of Machine Learning for Aboveground Biomass Modeling in Tropical and Temperate Forests from Airborne Hyperspectral Imagery
- Use What You Know: Causal Foundation Models with Partial Graphs
- The steep cost of capture
- Monitoring urban construction and quarry blasts with low-cost seismic sensors and deep learning tools in the city of Oslo, Norway
- Identifying controlling factors of delta morphology using a convolutional autoencoder
- The Artificial Intelligence Cognitive Examination: A Survey on the Evolution of Multimodal Evaluation From Recognition to Reasoning
- EPick: Multi-Class Attention-based U-shaped Neural Network for Earthquake Detection and Seismic Phase Picking
- A Computer-Aided Diagnosis System for Breast Pathology: A Deep Learning Approach with Model Interpretability from Pathological Perspective
- A Performance Comparison of Loss Functions for Deep Face Recognition
- A highly selective response to food in human visual cortex revealed by hypothesis-free voxel decomposition
- Feature Space Augmentation for Long-Tailed Data
- On the Generalization Error Bounds of Neural Networks under Diversity-Inducing Mutual Angular Regularization
- Wireless for Machine Learning
- Auto-Encoding Twin-Bottleneck Hashing
- End-to-End Learning Local Multi-view Descriptors for 3D Point Clouds
- A Geometric Approach to Online Streaming Feature Selection
- Deep Affinity Net: Instance Segmentation via Affinity
- Learn to Augment: Joint Data Augmentation and Network Optimization for Text Recognition
- Signal-to-Noise Ratio: A Robust Distance Metric for Deep Metric Learning
- GRIPHIN: grids of pharmacophore interaction fields for affinity prediction
- The Brain Abstracted
- Edge-Enhanced Vision Transformer Framework for Accurate AI-Generated Image Detection
- Overcoming Data Sparsity in Group Recommendation
- Learning from Large-scale Noisy Web Data with Ubiquitous Reweighting for Image Classification
- Not All Ops Are Created Equal!
- Handcrafted Backdoors in Deep Neural Networks
- A Survey of Autonomous Driving: Common Practices and Emerging Technologies
- What Do Single-view 3D Reconstruction Networks Learn?
- AutoFlow: Learning a Better Training Set for Optical Flow
- Algebraic Approach to Ridge-Regularized Mean Squared Error Minimization in Minimal ReLU Neural Network
- Identifying and Compensating for Feature Deviation in Imbalanced Deep Learning
- Merchant Category Identification Using Credit Card Transactions
- Shape and Symmetry Induction for 3D Objects
- Detection of Furigana Text in Images
- Game-Theoretic Multiagent Reinforcement Learning
- Can We Faithfully Represent Masked States to Compute Shapley Values on a DNN?
- Robotic grasp detection using a novel two-stage approach
- TRACE: Early Detection of Chronic Kidney Disease Onset with Transformer-Enhanced Feature Embedding
- LoCo: Local Contrastive Representation Learning
- Biologically-inspired Salience Affected Artificial Neural Network (SANN)
- Video Anomaly Detection by Estimating Likelihood of Representations
- A Sheaf and Topology Approach to Generating Local Branch Numbers in Digital Images
- Robustness and Transferability of Universal Attacks on Compressed Models
- Seeing eye-to-eye? A comparison of object recognition performance in humans and deep convolutional neural networks under image manipulation
- Uncertainty-driven ensembles of deep architectures for multiclass classification. Application to COVID-19 diagnosis in chest X-ray images
- Why Convolutional Networks Learn Oriented Bandpass Filters: Theory and Empirical Support
- DSRNA: Differentiable Search of Robust Neural Architectures
- ParaNet: Deep Regular Representation for 3D Point Clouds
- KVL-BERT: Knowledge Enhanced Visual-and-Linguistic BERT for Visual Commonsense Reasoning
- Meticulous Object Segmentation
- Quantum Computing Tools for Fast Detection of Gravitational Waves in the Context of LISA Space Mission
- Sketch-Specific Data Augmentation for Freehand Sketch Recognition
- Measurement-driven Security Analysis of Imperceptible Impersonation Attacks
- The Sloop System for Individual Animal Identification with Deep Learning
- Assessing the Alignment of Popular CNNs to the Brain for Valence Appraisal
- ROPA: Synthetic Robot Pose Generation for RGB-D Bimanual Data Augmentation
- Emergent symbolic language based deep medical image classification
- Pure Vision Language Action (VLA) Models: A Comprehensive Survey
- Visual Recognition Using Directional Distribution Distance
- ViG-LRGC: Vision Graph Neural Networks with Learnable Reparameterized Graph Construction
- Improving Outdoor Multi-cell Fingerprinting-based Positioning via Mobile Data Augmentation
- Hardness-Aware Deep Metric Learning
- Online and Offline Handwritten Chinese Character Recognition: A Comprehensive Study and New Benchmark
- Freehand Sketch Recognition Using Deep Features
- A Tutorial on Quantum Convolutional Neural Networks (QCNN)
- Uncertainty-aware Short-term Motion Prediction of Traffic Actors for Autonomous Driving
- LEAF-Mamba: Local Emphatic and Adaptive Fusion State Space Model for RGB-D Salient Object Detection
- Density-embedding layers: a general framework for adaptive receptive fields
- A Scalable Lift-and-Project Differentiable Approach For the Maximum Cut Problem
- OmniFed: A Modular Framework for Configurable Federated Learning from Edge to HPC
- Machine learning approach to single-shot multiparameter estimation for the non-linear Schrödinger equation
- Automated Tracking of Primate Behavior
- Bounded PCTL Model Checking of Large Language Model Outputs
- Recognition Of Surface Defects On Steel Sheet Using Transfer Learning
- Detecting Deep Neural Network Defects with Data Flow Analysis
- Self-attention based BiLSTM-CNN classifier for the prediction of ischemic and non-ischemic cardiomyopathy
- Effective and efficient ROI-wise visual encoding using an end-to-end CNN regression model and selective optimization
- Towards Understanding and Modeling Empathy for Use in Motivational Design Thinking
- A Deep Ordinal Distortion Estimation Approach for Distortion Rectification
- SSFN -- Self Size-estimating Feed-forward Network with Low Complexity, Limited Need for Human Intervention, and Consistent Behaviour across Trials
- Six Sigma For Neural Networks: Taguchi-based optimization
- Learning Invariant Representations for Sentiment Analysis: The Missing Material is Datasets
- Effectiveness of Data Augmentation in Cellular-based Localization Using Deep Learning
- DINOv3-Diffusion Policy: Self-Supervised Large Visual Model for Visuomotor Diffusion Policy Learning
- From Benchmarks to Reality: Advancing Visual Anomaly Detection by the VAND 3.0 Challenge
- Modeling Human Motion with Quaternion-based Neural Networks
- The Evolution of Neural Network-Based Chart Patterns: A Preliminary Study
- Neural network models and deep learning - a primer for biologists
- FPGA-based Accelerators of Deep Learning Networks for Learning and Classification: A Review
- Trustworthy Convolutional Neural Networks: A Gradient Penalized-based Approach
- A PMU-Based Machine Learning Application for Fast Detection of Forced Oscillations from Wind Farms
- Automatic Intermodal Loading Unit Identification using Computer Vision: A Scoping Review
- nDNA -- the Semantic Helix of Artificial Cognition
- SynergyNet: Fusing Generative Priors and State-Space Models for Facial Beauty Prediction
- SnipSnap: A Joint Compression Format and Dataflow Co-Optimization Framework for Efficient Sparse LLM Accelerator Design
- From handcrafted to deep local features
- Deep Clustering for Unsupervised Learning of Visual Features
- Guidelines and Benchmarks for Deployment of Deep Learning Models on\n Smartphones as Real-Time Apps
- A transfer learning method with deep residual network for pediatric pneumonia diagnosis
- Neural Network Based Framework for Passive Intermodulation Cancellation in MIMO Systems
- Bit-Flip Attack: Crushing Neural Network with Progressive Bit Search
- A multiscale neural network based on hierarchical nested bases
- SOLAR: Switchable Output Layer for Accuracy and Robustness in Once-for-All Training
- Towards a Transparent and Interpretable AI Model for Medical Image Classifications
- Checking extracted rules in Neural Networks
- Randomized Smoothing Meets Vision-Language Models
- Ensemble Distillation for Robust Model Fusion in Federated Learning
- Deep learning algorithms for detection of diabetic retinopathy in retinal fundus photographs: A systematic review and meta-analysis
- Skeletal bone age prediction based on a deep residual network with spatial transformer
- A neural network approach to segment brain blood vessels in digital subtraction angiography
- ChartMaster: Advancing Chart-to-Code Generation with Real-World Charts and Chart Similarity Reinforcement Learning
- A vision transformer based approach for analysis of plasmodium vivax life cycle for malaria prediction using thin blood smear microscopic images
- A deep sift convolutional neural networks for total brain volume estimation from 3D ultrasound images
- Deep Learning and Traffic Classification: Lessons learned from a commercial-grade dataset with hundreds of encrypted and zero-day applications
- Region-Aware Deformable Convolutions
- Recent Advancements in Microscopy Image Enhancement using Deep Learning: A Survey
- Effect of the initial configuration of weights on the training and function of artificial neural networks
- Dissecting the Graphcore IPU Architecture via Microbenchmarking
- RGB-Only Supervised Camera Parameter Optimization in Dynamic Scenes
- Model Compression and Hardware Acceleration for Neural Networks: A Comprehensive Survey
- Text Classification Improved by Integrating Bidirectional LSTM with Two-dimensional Max Pooling
- Crafting GBD-Net for Object Detection
- Structure-Preserving Margin Distribution Learning for High-Order Tensor Data with Low-Rank Decomposition
- Exploring Self-attention for Image Recognition
- Task-Aware Monocular Depth Estimation for 3D Object Detection
- White-box machine learning for uncovering physically interpretable dimensionless governing equations for granular materials
- Adaptable image quality assessment using meta-reinforcement learning of task amenability
- Domain Generalization via Multidomain Discriminant Analysis
- Pointing Novel Objects in Image Captioning
- Survey on semantic segmentation using deep learning techniques
- Deep Learning For Computer Vision Tasks: A review
- Open-World Class Discovery with Kernel Networks
- Dynamic Spectrum Matching with One-shot Learning
- Unsupervised Object-Level Representation Learning from Scene Images
- MARIC: Multi-Agent Reasoning for Image Classification
- Learning FRAME Models Using CNN Filters
- Incorporating Visual Cortical Lateral Connection Properties into CNN: Recurrent Activation and Excitatory-Inhibitory Separation
- Deep Predictive Coding Network for Object Recognition
- Probabilistic Neural Network with Complex Exponential Activation Functions in Image Recognition using Deep Learning Framework
- Techniques for visualizing LSTMs applied to electrocardiograms
- Can you hear me now? Sensitive comparisons of human and machine perception
- DeepPeep: Exploiting Design Ramifications to Decipher the Architecture of Compact DNNs
- Towards Robust Defense against Customization via Protective Perturbation Resistant to Diffusion-based Purification
- Binary Classification of Light and Dark Time Traces of a Transition Edge Sensor Using Convolutional Neural Networks
- Embedding Label Structures for Fine-Grained Feature Representation
- Bridging the Gap between Label- and Reference-based Synthesis in Multi-attribute Image-to-Image Translation
- Rest2Visual: Predicting Visually Evoked fMRI from Resting-State Scans
- A Lightweight ReLU-Based Feature Fusion for Aerial Scene Classification
- TreeIRL: Safe Urban Driving with Tree Search and Inverse Reinforcement Learning
- Intelligent Vacuum Thermoforming Process
- RTMobile: Beyond Real-Time Mobile Acceleration of RNNs for Speech Recognition
- Robot-Assisted Feeding: Generalizing Skewering Strategies across Food Items on a Realistic Plate
- On the Structural Sensitivity of Deep Convolutional Networks to the Directions of Fourier Basis Functions
- Deep Learning for Design and Retrieval of Nano-photonic Structures
- Transfer Metric Learning: Algorithms, Applications and Outlooks
- Proof of Federated Learning: A Novel Energy-recycling Consensus Algorithm
- EByFTVeS: Efficient Byzantine Fault Tolerant-based Verifiable Secret-sharing in Distributed Privacy-preserving Machine Learning
- TreeGAN: Syntax-Aware Sequence Generation with Generative Adversarial Networks
- Revisiting Fine-tuning for Few-shot Learning
- BATR-FST: Bi-Level Adaptive Token Refinement for Few-Shot Transformers
- Multi-level Wavelet Convolutional Neural Networks
- High-resolution rainfall-runoff modeling using graph neural network
- DARTS: Deceiving Autonomous Cars with Toxic Signs
- A Robust and Precise ConvNet for small non-coding RNA classification\n (RPC-snRC)
- A Data-Aware Fourier Neural Operator for Modeling Spatiotemporal Electromagnetic Fields
- Ranking Distillation: Learning Compact Ranking Models With High Performance for Recommender System
- Learning Low-rank Deep Neural Networks via Singular Vector Orthogonality Regularization and Singular Value Sparsification
- A Survey on Neural Architecture Search
- On the Decision Boundary of Deep Neural Networks
- A Loss Function for Generative Neural Networks Based on Watson's Perceptual Model
- Spherical Convolutional Neural Networks: Stability to Perturbations in SO(3)
- Advancing chest X-ray diagnostics: A novel CycleGAN-based preprocessing approach for enhanced lung disease classification in ChestX-Ray14
- Proposal Learning for Semi-Supervised Object Detection
- A Random Matrix Perspective on Mixtures of Nonlinearities for Deep Learning
- Overlearning Reveals Sensitive Attributes
- Toward Filament Segmentation Using Deep Neural Networks
- Microsoft COCO Captions: Data Collection and Evaluation Server
- Future Frame Prediction for Anomaly Detection -- A New Baseline
- Tubelets: Unsupervised action proposals from spatiotemporal super-voxels
- Scene Labeling with Contextual Hierarchical Models
- VATT: Transformers for Multimodal Self-Supervised Learning from Raw Video, Audio and Text
- FusionMAE: large-scale pretrained model to optimize and simplify diagnostic and control of fusion plasma
- PhotoApp: Photorealistic Appearance Editing of Head Portraits
- Re-labeling ImageNet: from Single to Multi-Labels, from Global to Localized Labels
- Beyond Triplet Loss: Meta Prototypical N-tuple Loss for Person Re-identification
- RGP: Neural Network Pruning through Its Regular Graph Structure
- TFD-former: Time-frequency domain fusion decoders for effective and robust fault diagnosis under time-varying speeds
- SCA-PVNet: Self-and-cross attention based aggregation of point cloud and multi-view for 3D object retrieval
- Subject-independent Human Pose Image Construction with Commodity Wi-Fi
- MailLeak: Obfuscation-Robust Character Extraction Using Transfer Learning
- Relating Graph Neural Networks to Structural Causal Models
- Deep Learning in Memristive Nanowire Networks
- Univariate ReLU neural network and its application in nonlinear system identification
- Learning Two-Branch Neural Networks for Image-Text Matching Tasks
- Uncovering the Limits of Adversarial Training against Norm-Bounded Adversarial Examples
- Classifier Crafting: Turn Your ConvNet into a Zero-Shot Learner!
- Learning Deep Representations of Fine-grained Visual Descriptions
- Addressing Failure Prediction by Learning Model Confidence
- Flows Over Periodic Hills of Parameterized Geometries: A Dataset for Data-Driven Turbulence Modeling From Direct Simulations
- Autonomous Driving with Deep Learning: A Survey of State-of-Art Technologies
- MaX-DeepLab: End-to-End Panoptic Segmentation with Mask Transformers
- Causal-Symbolic Meta-Learning (CSML): Inducing Causal World Models for Few-Shot Generalization
- Synetgy: Algorithm-hardware Co-design for ConvNet Accelerators on Embedded FPGAs
- Deep learning with asymmetric connections and Hebbian updates
- Sitatapatra: Blocking the Transfer of Adversarial Samples
- A survey on deep learning approaches for breast cancer diagnosis
- Data augmentation and image understanding
- Artificial neural networks for neuroscientists: A primer
- LEGO: Spatial Accelerator Generation and Optimization for Tensor Applications
- Normalization Techniques in Training DNNs: Methodology, Analysis and Application
- Augmented Skeleton Based Contrastive Action Learning with Momentum LSTM for Unsupervised Action Recognition
- Unrolling Graph-based Douglas-Rachford Algorithm for Image Interpolation with Informed Initialization
- A Multiplexed Network for End-to-End, Multilingual OCR
- SugarcaneShuffleNet: A Very Fast, Lightweight Convolutional Neural Network for Diagnosis of 15 Sugarcane Leaf Diseases
- Customizing Student Networks From Heterogeneous Teachers via Adaptive Knowledge Amalgamation
- Efficient Algorithms for Device Placement of DNN Graph Operators
- The Neural Network Approach to Inverse Problems in Differential Equations
- Evaluating State-of-the-Art Classification Models Against Bayes Optimality
- IS-Diff: Improving Diffusion-Based Inpainting with Better Initial Seed
- Determining the boundary of dynamical chaos in the generalized Chirikov map via machine learning
- HyperInverter: Improving StyleGAN Inversion via Hypernetwork
- Recognizing American Sign Language Manual Signs from RGB-D Videos
- Intention-aware Long Horizon Trajectory Prediction of Surrounding Vehicles using Dual LSTM Networks
- A Controllable 3D Deepfake Generation Framework with Gaussian Splatting
- Efficient Byzantine-Robust Privacy-Preserving Federated Learning via Dimension Compression
- Optimizing Latent Dimension Allocation in Hierarchical VAEs: Balancing Attenuation and Information Retention for OOD Detection
- DISA at ImageCLEF 2014 Revised: Search-based Image Annotation with DeCAF Features
- Error Control and Loss Functions for the Deep Learning Inversion of Borehole Resistivity Measurements
- Mitigate Parasitic Resistance in Resistive Crossbar-based Convolutional Neural Networks
- gen2Out: Detecting and Ranking Generalized Anomalies
- Enhancing Electromagnetic Calorimeter Signal Reconstruction with Machine Learning-Based Noise Discrimination
- AngularGrad: A New Optimization Technique for Angular Convergence of Convolutional Neural Networks
- Straggler-resistant distributed matrix computation via coding theory
- InterBERT: Vision-and-Language Interaction for Multi-modal Pretraining
- Neural networks in the search for fast radio bursts with RATAN-600
- Generalization Bounds of Stochastic Gradient Descent for Wide and Deep Neural Networks
- Self-Attention Capsule Networks for Object Classification
- Squeeze-and-Excitation on Spatial and Temporal Deep Feature Space for Action Recognition
- An Advanced Convolutional Neural Network for Bearing Fault Diagnosis under Limited Data
- Compact Device Models for FinFET and Beyond
- Unlabeled Data Deployment for Classification of Diabetic Retinopathy Images Using Knowledge Transfer
- Weakly Supervised Vulnerability Localization via Multiple Instance Learning
- Sinogram super-resolution and denoising convolutional neural network (SRCN) for limited data photoacoustic tomography
- Review: deep learning on 3D point clouds
- Two-Phase Object-Based Deep Learning for Multi-temporal SAR Image Change Detection
- A survey on Deep Learning based bearing fault diagnosis
- Neural Autoregressive Distribution Estimation
- DeepTraverse: A Depth-First Search Inspired Network for Algorithmic Visual Understanding
- Learnable Gabor modulated complex-valued networks for orientation robustness
- Deep Learning for Generic Object Detection: A Survey
- Rethinking Channel Dimensions for Efficient Model Design
- Characterizing the Efficiency of Distributed Training: A Power, Performance, and Thermal Perspective
- Deep Learning for Insider Threat Detection: Review, Challenges and Opportunities
- A Critical Review of Recurrent Neural Networks for Sequence Learning
- Identification of Crystal Symmetry from Noisy Diffraction Patterns by A Shape Analysis and Deep Learning
- A Sketch Based 3D Shape Retrieval Approach Based on Efficient Deep Point-to-Subspace Metric Learning
- A Capsule-unified Framework of Deep Neural Networks for Graphical Programming
- Knowledge Transfer via Dense Cross-Layer Mutual-Distillation
- Deep Convolutional Neural Networks in the Face of Caricature: Identity and Image Revealed
- Exploring Weight Symmetry in Deep Neural Networks
- InverSynth: Deep Estimation of Synthesizer Parameter Configurations from Audio Signals
- Contrastive Learning with Stronger Augmentations
- Dense RepPoints: Representing Visual Objects with Dense Point Sets
- Identifying Pediatric Vascular Anomalies With Deep Learning
- Parameterized Knowledge Transfer for Personalized Federated Learning
- STFCN: Spatio-Temporal FCN for Semantic Video Segmentation
- NAT: Learning to Attack Neurons for Enhanced Adversarial Transferability
- Expressive Power of Deep Networks on Manifolds: Simultaneous Approximation
- Person Re-identification: Past, Present and Future
- Image Deformation Meta-Networks for One-Shot Learning
- Objectness Similarity: Capturing Object-Level Fidelity in 3D Scene Evaluation
- Illumination-aware Faster R-CNN for Robust Multispectral Pedestrian Detection
- Automatic Cross-Replica Sharding of Weight Update in Data-Parallel Training
- MSDANet: A Multiscale Dual-Channel Spatial Attention Network with Depthwise Separable Convolution for Hyperspectral Image Classification
- Images in Motion?: A First Look into Video Leakage in Collaborative Deep Learning
- MultiGrain: a unified image embedding for classes and instances
- Diffusion-Based Action Recognition Generalizes to Untrained Domains
- Deep Learning for LiDAR Point Clouds in Autonomous Driving: A Review
- Renovating Parsing R-CNN for Accurate Multiple Human Parsing
- Compressing CNN models for resource-constrained systems by channel and layer pruning
- Modulating human brain responses via optimal natural image selection and synthetic image generation
- CNN-ViT Hybrid for Pneumonia Detection: Theory and Empiric on Limited Data without Pretraining
- Network Pruning via Transformable Architecture Search
- Deep Learning for Free-Hand Sketch: A Survey
- Line Segment Detection Using Transformers without Edges
- Multispectral CT Denoising via Simulation-Trained Deep Learning: Experimental Results at the ESRF BM18
- Review of Video Predictive Understanding: Early Action Recognition and Future Action Prediction
- Actor-Centric Relation Network
- EAST: An Efficient and Accurate Scene Text Detector
- Trellis Networks for Sequence Modeling
- Distilling the Knowledge in a Neural Network
- Boosted Training of Lightweight Early Exits for Optimizing CNN Image Classification Inference
- Dual-Thresholding Heatmaps to Cluster Proposals for Weakly Supervised Object Detection
- Label Smoothing++: Enhanced Label Regularization for Training Neural Networks
- RISC-NN: Use RISC, NOT CISC as Neural Network Hardware Infrastructure
- Rollout-LaSDI: Enhancing the long-term accuracy of Latent Space Dynamics
- Localized PCA-Net Neural Operators for Scalable Solution Reconstruction of Elliptic PDEs
- Impression Space from Deep Template Network
- Residual Dense Network for Image Restoration
- A Practical Deep Learning-Based Acoustic Side Channel Attack on Keyboards
- FusionLane: Multi-Sensor Fusion for Lane Marking Semantic Segmentation Using Deep Neural Networks
- PBRnet: Pyramidal Bounding Box Refinement to Improve Object Localization Accuracy
- Efficient resource management in UAVs for Visual Assistance
- Dataset Distillation
- Language Self-Play For Data-Free Training
- Three Pillars improving Vision Foundation Model Distillation for Lidar
- Spectral and Rhythm Feature Performance Evaluation for Category and Class Level Audio Classification with Deep Convolutional Neural Networks
- GLEAM: Learning to Match and Explain in Cross-View Geo-Localization
- Temporal Image Forensics: A Review and Critical Evaluation
- Enhanced Memory Network: The novel network structure for Symbolic Music Generation
- Revisiting the Calibration of Modern Neural Networks
- Stochastic Sign Descent Methods: New Algorithms and Better Theory
- Inferring brain-computational mechanisms with models of activity measurements
- Evaluating the Impact of Adversarial Attacks on Traffic Sign Classification using the LISA Dataset
- Data-driven discovery of dynamical models in biology
- Video Anomaly Detection and Localization via Gaussian Mixture Fully Convolutional Variational Autoencoder
- An Analysis of Scale Invariance in Object Detection - SNIP
- Expert-Guided Explainable Few-Shot Learning for Medical Image Diagnosis
- Foldover Features for Dynamic Object Behavior Description in Microscopic Videos
- Deep learning visual analysis in laparoscopic surgery: a systematic review and diagnostic test accuracy meta-analysis
- Comparison Network for One-Shot Conditional Object Detection
- Machine Learning: Algorithms, Real-World Applications and Research Directions
- Unifying Remote Sensing Image Retrieval and Classification with Robust Fine-tuning
- Adaptive, Distribution-Free Prediction Intervals for Deep Networks
- Laguerre-Gauss Preprocessing: Line Profiles as Image Features for Aerial Images Classification
- Image Enhanced Rotation Prediction for Self-Supervised Learning
- Dimensionally Reduced Open-World Clustering: DROWCULA
- Lookup multivariate Kolmogorov-Arnold Networks
- AIM 2025 Challenge on High FPS Motion Deblurring: Methods and Results
- Albumentations: fast and flexible image augmentations
- Data-Efficient Time-Dependent PDE Surrogates: Graph Neural Simulators vs. Neural Operators
- RPC: A Large-Scale Retail Product Checkout Dataset
- Just Jump: Dynamic Neighborhood Aggregation in Graph Neural Networks
- Fooling Computer Vision into Inferring the Wrong Body Mass Index
- Micro-Expression Recognition via Fine-Grained Dynamic Perception
- Agglomerative Attention
- Three-Dimensional Mesh Steganography and Steganalysis: A Review
- A brain-inspired paradigm for scalable quantum vision
- Unity Style Transfer for Person Re-Identification
- DNA: Deeply-supervised Nonlinear Aggregation for Salient Object Detection
- DeepPoison: Feature Transfer Based Stealthy Poisoning Attack
- Explainable AI: A Review of Machine Learning Interpretability Methods
- Dual-Mode Deep Anomaly Detection for Medical Manufacturing: Structural Similarity and Feature Distance
- High Utilization Energy-Aware Real-Time Inference Deep Convolutional Neural Network Accelerator
- Simulation Priors for Data-Efficient Deep Learning
- Rethinking Supervised Pre-training for Better Downstream Transferring
- Prior Distribution and Model Confidence
- Pipe-SGD: A Decentralized Pipelined SGD Framework for Distributed Deep Net Training
- Systematic Review and Meta-analysis of AI-driven MRI Motion Artifact Detection and Correction
- Lite-HRNet: A Lightweight High-Resolution Network
- Space-time Mixing Attention for Video Transformer
- TemporalFlowViz: Parameter-Aware Visual Analytics for Interpreting Scramjet Combustion Evolution
- Combating the Elsagate phenomenon: Deep learning architectures for disturbing cartoons
- Synergy: Resource Sensitive DNN Scheduling in Multi-Tenant Clusters
- Graph Unlearning: Efficient Node Removal in Graph Neural Networks
- Advanced Brain Tumor Segmentation Using EMCAD: Efficient Multi-scale Convolutional Attention Decoding
- Scale-interaction transformer: a hybrid cnn-transformer model for facial beauty prediction
- Deep Learning Approaches for Image Retrieval and Pattern Spotting in Ancient Documents
- Multi-Class Lane Semantic Segmentation using Efficient Convolutional Networks
- Convolutional Dictionary Learning in Hierarchical Networks
- Dynamic Sensitivity Filter Pruning using Multi-Agent Reinforcement Learning For DCNN's
- On the Normalization of Confusion Matrices: Methods and Geometric Interpretations
- Bayesian Loss for Crowd Count Estimation with Point Supervision
- SpecNet: Spectral Domain Convolutional Neural Network
- Universality of Gradient Descent Neural Network Training
- Video Analytics with Zero-streaming Cameras
- Towards Open World Detection: A Survey
- Empirical Studies on the Properties of Linear Regions in Deep Neural Networks
- VCMamba: Bridging Convolutions with Multi-Directional Mamba for Efficient Visual Representation
- Measuring the Measures: Discriminative Capacity of Representational Similarity Metrics Across Model Families
- Delta Activations: A Representation for Finetuned Large Language Models
- Hyper Diffusion Avatars: Dynamic Human Avatar Generation using Network Weight Space Diffusion
- Light Field Reconstruction via Deep Adaptive Fusion of Hybrid Lenses
- Differential Morphological Profile Neural Networks for Semantic Segmentation
- Transformer-based Multi-Aspect Modeling for Multi-Aspect Multi-Sentiment Analysis
- A Survey on 3D Skeleton-Based Action Recognition Using Learning Method
- Lesion-based Contrastive Learning for Diabetic Retinopathy Grading from Fundus Images
- Improving the Accuracy and Hardware Efficiency of Neural Networks Using Approximate Multipliers
- Interpreting Black-Box Models: A Review on Explainable Artificial Intelligence
- SGPN: Similarity Group Proposal Network for 3D Point Cloud Instance Segmentation
- Deep Gamblers: Learning to Abstain with Portfolio Theory
- Understanding Architectures Learnt by Cell-based Neural Architecture Search
- 1D Convolutional Neural Networks and Applications: A Survey
- From Predictions to Explanations: Explainable AI for Autism Diagnosis and Identification of Critical Brain Regions
- Spatiotemporal Pyramid Network for Video Action Recognition
- ResiliNet: Failure-Resilient Inference in Distributed Neural Networks
- How Important is the Train-Validation Split in Meta-Learning?
- Watermarking Graph Neural Networks based on Backdoor Attacks
- Semantic Segmentation from Limited Training Data
- Moment Matching for Multi-Source Domain Adaptation
- Robust Training of Social Media Image Classification Models for Rapid Disaster Response
- Learning to See before Learning to Act: Visual Pre-training for Manipulation
- Sparse Autoencoder Neural Operators: Model Recovery in Function Spaces
- LSAM: Asynchronous Distributed Training with Landscape-Smoothed Sharpness-Aware Minimization
- FastCaps: A Design Methodology for Accelerating Capsule Network on Field Programmable Gate Arrays
- Isolated Bangla Handwritten Character Classification using Transfer Learning
- Minimizing Perceived Image Quality Loss Through Adversarial Attack Scoping
- CSWin Transformer: A General Vision Transformer Backbone with Cross-Shaped Windows
- Understanding deep learning requires rethinking generalization
- End-to-end acoustic modeling using convolutional neural networks for HMM-based automatic speech recognition
- Directional Bias Amplification
- Long-Term Vehicle Localization by Recursive Knowledge Distillation
- Network Implosion: Effective Model Compression for ResNets via Static Layer Pruning and Retraining
- Human De-occlusion: Invisible Perception and Recovery for Humans
- MIDI-Sandwich2: RNN-based Hierarchical Multi-modal Fusion Generation VAE networks for multi-track symbolic music generation
- Residual Squeeze VGG16
- Stealth by Conformity: Evading Robust Aggregation through Adaptive Poisoning
- Multi-Scale Deep Learning for Colon Histopathology: A Hybrid Graph-Transformer Approach
- A Polynomial-Based Approach for Architectural Design and Learning with\n Deep Neural Networks
- Making EfficientNet More Efficient: Exploring Batch-Independent Normalization, Group Convolutions and Reduced Resolution Training
- A Convolutional Hierarchical Deep-learning Neural Network (C-HiDeNN) Framework for Non-linear Finite Element Analysis
- Vision encoders should be image size agnostic and task driven
- MixKD: Towards Efficient Distillation of Large-scale Language Models
- Vision-based deep execution monitoring
- MobileFaceNets: Efficient CNNs for Accurate Real-Time Face Verification on Mobile Devices
- Exploring Vicinal Risk Minimization for Lightweight Out-of-Distribution Detection
- The gap between theory and practice in function approximation with deep neural networks
- On the Generalization of Models Trained with SGD: Information-Theoretic Bounds and Implications
- Open-Domain Conversational Agents: Current Progress, Open Problems, and Future Directions
- FAWA: Fast Adversarial Watermark Attack on Optical Character Recognition (OCR) Systems
- Pruning Convolutional Neural Networks with Self-Supervision
- Robots of the Lost Arc: Self-Supervised Learning to Dynamically Manipulate Fixed-Endpoint Cables
- Robust Small Methane Plume Segmentation in Satellite Imagery
- One Size Does Not Fit All: Multi-Scale, Cascaded RNNs for Radar Classification
- LaSOT: A High-quality Large-scale Single Object Tracking Benchmark
- TensorLib: A Spatial Accelerator Generation Framework for Tensor Algebra
- HourNAS: Extremely Fast Neural Architecture Search Through an Hourglass Lens
- An Investigation of Visual Foundation Models Robustness
- Social gaze fingerprints: identifying social virtual reality users by their eye gaze patterns
- Learnable Loss Geometries with Mirror Descent for Scalable and Convergent Meta-Learning
- ViTAE: Vision Transformer Advanced by Exploring Intrinsic Inductive Bias
- Climate impacts and future trends of hailstorms in China based on millennial records
- Models Matter, So Does Training: An Empirical Study of CNNs for Optical Flow Estimation
- Synesthesia of Machines (SoM)-Based Task-Driven MIMO System for Image Transmission
- Characterizing Types of Convolution in Deep Convolutional Recurrent Neural Networks for Robust Speech Emotion Recognition
- Generalized Zero-Shot Domain Adaptation via Coupled Conditional Variational Autoencoders
- Memory Efficient Class-Incremental Learning for Image Classification
- URLNet: Learning a URL Representation with Deep Learning for Malicious URL Detection
- Reinforcement Learning for Weakly Supervised Temporal Grounding of Natural Language in Untrimmed Videos
- Rotation Invariance Neural Network
- Practical Detection of Trojan Neural Networks: Data-Limited and Data-Free Cases
- Diversity inducing Information Bottleneck in Model Ensembles
- A Runtime-Based Computational Performance Predictor for Deep Neural Network Training
- Using Capsule Neural Network to predict Tuberculosis in lens-free microscopic images
- MemeSequencer: Sparse Matching for Embedding Image Macros
- Elastic Consistency: A General Consistency Model for Distributed Stochastic Gradient Descent
- Music Genre Classification Using Machine Learning Techniques
- TransMatch: A Transfer-Learning Framework for Defect Detection in Laser Powder Bed Fusion Additive Manufacturing
- SoK: Understanding the Fundamentals and Implications of Sensor Out-of-band Vulnerabilities
- Mamba-CNN: A Hybrid Architecture for Efficient and Accurate Facial Beauty Prediction
- Lightweight and Fast Real-time Image Enhancement via Decomposition of the Spatial-aware Lookup Tables
- Towards More Diverse and Challenging Pre-training for Point Cloud Learning: Self-Supervised Cross Reconstruction with Decoupled Views
- Contrastive Model Inversion for Data-Free Knowledge Distillation
- MSE Loss with Outlying Label for Imbalanced Classification
- Learning Connectivity of Neural Networks from a Topological Perspective
- Towards Deep Learning Assisted Autonomous UAVs for Manipulation Tasks in GPS-Denied Environments
- Plant Disease Detection and Classification by Deep Learning—A Review
- Review on Convolutional Neural Network (CNN) Applied to Plant Leaf Disease Classification
- Proximal Backpropagation
- SocialGCN: An Efficient Graph Convolutional Network based Model for Social Recommendation
- Performance evaluation of an integrated photonic convolutional neural network based on delay buffering and wavelength division multiplexing
- The Limitations of Deep Learning in Adversarial Settings
- Move Evaluation in Go Using Deep Convolutional Neural Networks
- Probabilistic Permutation Synchronization using the Riemannian Structure\n of the Birkhoff Polytope
- OnlineAugment: Online Data Augmentation with Less Domain Knowledge
- Assessing the (Un)Trustworthiness of Saliency Maps for Localizing Abnormalities in Medical Imaging
- Routing Towards Discriminative Power of Class Capsules
- From Synthetic to Real: Unsupervised Domain Adaptation for Animal Pose Estimation
- COVID-19 personal protective equipment detection using real-time deep learning methods
- Modeling the Nonsmoothness of Modern Neural Networks
- Lung cancer identification: a review on detection and classification
- Compact Deep Aggregation for Set Retrieval
- Towards Accurate and Compact Architectures via Neural Architecture Transformer
- Understanding and Enhancing the Use of Context for Machine Translation
- Learning by training: emergent return-point memory from cyclically tuning disordered sphere packings
- Using Distance Estimation and Deep Learning to Simplify Calibration in Food Calorie Measurement
- Rethinking FUN: Frequency-Domain Utilization Networks
- Deep Learning on Image Denoising: An overview
- A differential neural network learns stochastic differential equations and the Black-Scholes equation for pricing multi-asset options
- Regularization via Mass Transportation
- Partial success in closing the gap between human and machine vision
- Application of discrete Ricci curvature in pruning randomly wired neural networks: A case study with chest x-ray classification of COVID-19
- C-DiffDet+: Fusing Global Scene Context with Generative Denoising for High-Fidelity Car Damage Detection
- Incremental Learning with Maximum Entropy Regularization: Rethinking Forgetting and Intransigence
- Unsupervised Domain Adaptation Learning Algorithm for RGB-D Staircase Recognition
- Dropout drops double descent
- CNN with large memory layers
- Generative Latent Space Dynamics of Electron Density
- Cross-Attention Multimodal Fusion for Breast Cancer Diagnosis: Integrating Mammography and Clinical Data with Explainability
- Understanding the Behaviour of Contrastive Loss
- Optimization Variance: Exploring Generalization Properties of DNNs
- Artificial Intelligence in Drug Discovery: Applications and Techniques
- AI Compute Architecture and Evolution Trends
- Representation Learning with Adaptive Superpixel Coding
- CuratorNet: Visually-aware Recommendation of Art Images
- Accelerated Training for Massive Classification via Dynamic Class Selection
- DeepOPF: Deep Neural Network for DC Optimal Power Flow
- OmniArt: Multi-task Deep Learning for Artistic Data Analysis
- Accelerating Federated Learning via Momentum Gradient Descent
- Convolutional Neural Fabrics
- Playing Atari with Deep Reinforcement Learning
- Survey of spiking in the mouse visual system reveals functional hierarchy
- EndoNet: A Deep Architecture for Recognition Tasks on Laparoscopic Videos
- Cyclone intensity estimate with context-aware cyclegan
- Deep Anomaly Detection by Residual Adaptation
- Estimating Model Uncertainty of Neural Networks in Sparse Information Form
- Collective Learning by Ensembles of Altruistic Diversifying Neural Networks
- Efficient Integer-Arithmetic-Only Convolutional Neural Networks
- Learning compact generalizable neural representations supporting perceptual grouping
- IA-RED2: Interpretability-Aware Redundancy Reduction for Vision Transformers
- GODS: Generalized One-class Discriminative Subspaces for Anomaly Detection
- Deep Verifier Networks: Verification of Deep Discriminative Models with Deep Generative Models
- Microscopic and collective signatures of feature learning in neural networks
- E-ConvNeXt: A Lightweight and Efficient ConvNeXt Variant with Cross-Stage Partial Connections
- Data-Free Adversarial Distillation
- Bridging Minds and Machines: Toward an Integration of AI and Cognitive Science
- D3PINNs: A Novel Physics-Informed Neural Network Framework for Staged Solving of Time-Dependent Partial Differential Equations
- Network Representation Learning: From Traditional Feature Learning to Deep Learning
- Developing a Multi-Modal Machine Learning Model For Predicting Performance of Automotive Hood Frames
- Understanding Incremental Learning with Closed-form Solution to Gradient Flow on Overparamerterized Matrix Factorization
- 3D-Aided Data Augmentation for Robust Face Understanding
- A Chain Graph Interpretation of Real-World Neural Networks
- Character-Aware Neural Language Models
- Improving Adversarial Robustness via Guided Complement Entropy
- Recovery Guarantees for Compressible Signals with Adversarial Noise
- Computational Emotion Analysis From Images: Recent Advances and Future Directions
- Iterative Low-Rank Approximation for CNN Compression
- What makes instance discrimination good for transfer learning?
- A Spatial-Frequency Aware Multi-Scale Fusion Network for Real-Time Deepfake Detection
- Exploring Randomly Wired Neural Networks for Image Recognition
- Objective Value Change and Shape-Based Accelerated Optimization for the Neural Network Approximation
- Seam360GS: Seamless 360° Gaussian Splatting from Real-World Omnidirectional Images
- Eigenvalue distribution of the Neural Tangent Kernel in the quadratic scaling
- Improving Calibration for Long-Tailed Recognition
- Radioactive data: tracing through training
- Constructing Geographic and Long-term Temporal Graph for Traffic Forecasting
- Measuring Dataset Granularity
- How Does Batch Normalization Help Optimization?
- Symplectic convolutional neural networks
- Improving Generalization in Deepfake Detection with Face Foundation Models and Metric Learning
- Task-Generalized Adaptive Cross-Domain Learning for Multimodal Image Fusion
- Causal Intervention for Weakly-Supervised Semantic Segmentation
- PRADA: Protecting against DNN Model Stealing Attacks
- Regularization With Stochastic Transformations and Perturbations for Deep Semi-Supervised Learning
- Multiscale Deep Equilibrium Models
- Mini-Batch Robustness Verification of Deep Neural Networks
- The GINN framework: a stochastic QED correspondence for stability and chaos in deep neural networks
- DeepOPF: A Deep Neural Network Approach for Security-Constrained DC Optimal Power Flow
- WIDER FACE: A Face Detection Benchmark
- Blended Coarse Gradient Descent for Full Quantization of Deep Neural Networks
- Alljoined-1.6M: A Million-Trial EEG-Image Dataset for Evaluating Affordable Brain-Computer Interfaces
- Return of the Devil in the Details: Delving Deep into Convolutional Nets
- Understanding the Effective Receptive Field in Deep Convolutional Neural Networks
- PanNuke Dataset Extension, Insights and Baselines
- Benchmarking Deep Reinforcement Learning for Continuous Control
- Bladder Cancer Diagnosis with Deep Learning: A Multi-Task Framework and Online Platform
- VQualA 2025 Challenge on Face Image Quality Assessment: Methods and Results
- Real-Time User-Guided Image Colorization with Learned Deep Priors
- Discretization-Aware Architecture Search
- PIPAL: a Large-Scale Image Quality Assessment Dataset for Perceptual Image Restoration
- Systematic evaluation of CNN advances on the ImageNet
- Robust and Efficient Quantum Reservoir Computing with Discrete Time Crystal
- Enhancing Adversarial Example Transferability with an Intermediate Level Attack
- AutoHOOT: Automatic High-Order Optimization for Tensors
- Use of Transfer Learning and Wavelet Transform for Breast Cancer Detection
- Ontology Based Global and Collective Motion Patterns for Event Classification in Basketball Videos
- Deep learning for the semi-classical limit of the Schrödinger equation
- Optimistic and Pessimistic Neural Networks for Scene and Object Recognition
- Attended End-to-end Architecture for Age Estimation from Facial Expression Videos
- ModelHub.AI: Dissemination Platform for Deep Learning Models
- Richer priors for infinitely wide multi-layer perceptrons
- Unpaired Image Translation via Adaptive Convolution-based Normalization
- Learning to Fuse Things and Stuff
- On the approximation of the solution of partial differential equations by artificial neural networks trained by a multilevel Levenberg-Marquardt method
- Frequency-adaptive tensor neural networks for high-dimensional multi-scale problems
- DesignCLIP: Multimodal Learning with CLIP for Design Patent Understanding
- Snap-Snap: Taking Two Images to Reconstruct 3D Human Gaussians in Milliseconds
- Study and development of a Computer-Aided Diagnosis system for classification of chest x-ray images using convolutional neural networks pre-trained for ImageNet and data augmentation
- Bias-based Universal Adversarial Patch Attack for Automatic Check-out
- Synthetic Adaptive Guided Embeddings (SAGE): A Novel Knowledge Distillation Method
- Real-time Human Detection Model for Edge Devices
- Convolutional Neural Network Pruning with Structural Redundancy Reduction
- Deep Learning for Taxol Exposure Analysis: A New Cell Image Dataset and Attention-Based Baseline Model
- On the Communication Latency of Wireless Decentralized Learning
- Minimizing Task-Oriented Age of Information for Remote Monitoring with Pre-Identification
- Improving OCR using internal document redundancy
- MoCHA-former: Moiré-Conditioned Hybrid Adaptive Transformer for Video Demoiréing
- Towards Precise End-to-end Weakly Supervised Object Detection Network
- Improved Mapping Between Illuminations and Sensors for RAW Images
- CP-mtML: Coupled Projection multi-task Metric Learning for Large Scale Face Retrieval
- Perceptual Adversarial Robustness: Defense Against Unseen Threat Models
- Neo: A Learned Query Optimizer
- Discriminative Noise Robust Sparse Orthogonal Label Regression-based Domain Adaptation
- Normalized Cut Loss for Weakly-supervised CNN Segmentation
- Analyzing Monotonic Linear Interpolation in Neural Network Loss Landscapes
- Lottery Jackpots Exist in Pre-trained Models
- Neural Image Beauty Predictor Based on Bradley-Terry Model
- Defeating Catastrophic Forgetting via Enhanced Orthogonal Weights Modification
- Ultra‐Low‐Field Paediatric MRI in Low‐ and Middle‐Income Countries: Super‐Resolution Using a Multi‐Orientation U‐Net
- Busy-Quiet Video Disentangling for Video Classification
- MOPS-Net: A Matrix Optimization-driven Network forTask-Oriented 3D Point Cloud Downsampling
- Capsule GAN Using Capsule Network for Generator Architecture
- MixHop: Higher-Order Graph Convolutional Architectures via Sparsified Neighborhood Mixing
- SVM and ELM: Who Wins? Object Recognition with Deep Convolutional Features from ImageNet
- On the Evaluation Metric for Hashing
- How to Train Your MAML to Excel in Few-Shot Classification
- torchgpipe: On-the-fly Pipeline Parallelism for Training Giant Models
- Representative Forgery Mining for Fake Face Detection
- CHAIN: Concept-harmonized Hierarchical Inference Interpretation of Deep Convolutional Neural Networks
- Data-Uncertainty Guided Multi-Phase Learning for Semi-Supervised Object Detection
- Open Set Domain Adaptation for Image and Action Recognition
- An Acoustic Segment Model Based Segment Unit Selection Approach to Acoustic Scene Classification with Partial Utterances
- A Survey on Video Anomaly Detection via Deep Learning: Human, Vehicle, and Environment
- Spherical U-Net on Cortical Surfaces: Methods and Applications
- One Shot vs. Iterative: Rethinking Pruning Strategies for Model Compression
- Towards Fewer Annotations: Active Learning via Region Impurity and Prediction Uncertainty for Domain Adaptive Semantic Segmentation
- FLAIR: Frequency- and Locality-Aware Implicit Neural Representations
- Evaluating Open-Source Vision Language Models for Facial Emotion Recognition against Traditional Deep Learning Models
- A Study of BFLOAT16 for Deep Learning Training
- Genuine multipartite entanglement verification with convolutional neural networks
- SpherePHD: Applying CNNs on a Spherical PolyHeDron Representation of 360 degree Images
- Training with the Invisibles: Obfuscating Images to Share Safely for Learning Visual Recognition Models
- Deeply Exploit Depth Information for Object Detection
- Lightweight Convolutional Representations for On-Device Natural Language Processing
- Unsupervised Domain Adaptation for Semantic Segmentation via Low-level Edge Information Transfer
- NeuroView: Explainable Deep Network Decision Making
- Optimization Problems for Machine Learning: A Survey
- ONG: One-Shot NMF-based Gradient Masking for Efficient Model Sparsification
- Uncertainty Propagation in Deep Neural Network Using Active Subspace
- D2-Mamba: Dual-Scale Fusion and Dual-Path Scanning with SSMs for Shadow Removal
- Creative4U: MLLMs-based Advertising Creative Image Selector with Comparative Reasoning
- SuryaBench: Benchmark Dataset for Advancing Machine Learning in Heliophysics and Space Weather Prediction
- Unsupervised Feature Learning by Cross-Level Instance-Group Discrimination
- The Pitfall of Evaluating Performance on Emerging AI Accelerators
- RED-NET: A Recursive Encoder-Decoder Network for Edge Detection
- NESTA: Hamming Weight Compression-Based Neural Proc. Engine
- Contextual Local Explanation for Black Box Classifiers
- Maximum-Entropy Adversarial Data Augmentation for Improved Generalization and Robustness
- Skin Cancer Classification: Hybrid CNN-Transformer Models with KAN-Based Fusion
- Efficient and Verifiable Privacy-Preserving Convolutional Computation for CNN Inference with Untrusted Clouds
- Attentive Deep Regression Networks for Real-Time Visual Face Tracking in Video Surveillance
- Seek and You Will Find: A New Optimized Framework for Efficient Detection of Pedestrian
- An ECC-based Fault Tolerance Approach for DNNs
- Joint Intent Detection and Slot Filling with Wheel-Graph Attention Networks
- Towards Generalizable Human Activity Recognition: A Survey
- One Policy to Control Them All: Shared Modular Policies for Agent-Agnostic Control
- Illusions in Humans and AI: How Visual Perception Aligns and Diverges
- Long Short-Term Relation Networks for Video Action Detection
- Belief-Conditioned One-Step Diffusion: Real-Time Trajectory Planning with Just-Enough Sensing
- TriQDef: Disrupting Semantic and Gradient Alignment to Prevent Adversarial Patch Transferability in Quantized Neural Networks
- ViPTT-Net: Video pretraining of spatio-temporal model for tuberculosis type classification from chest CT scans
- HALO: Learning to Prune Neural Networks with Shrinkage
- Importance Filtered Cross-Domain Adaptation
- Recent Advances and Applications of Deep Learning Methods in Materials Science
- Per-Tensor Fixed-Point Quantization of the Back-Propagation Algorithm
- A Comprehensive Review of AI Agents: Transforming Possibilities in Technology and Beyond
- End-to-End Learnable Geometric Vision by Backpropagating PnP Optimization
- PCA- and SVM-Grad-CAM for Convolutional Neural Networks: Closed-form Jacobian Expression
- The Rise of Generative AI for Metal-Organic Framework Design and Synthesis
- Itinerary-aware Personalized Deep Matching at Fliggy
- Weakly Supervised Estimation of Shadow Confidence Maps in Fetal Ultrasound Imaging
- DARB: A Density-Aware Regular-Block Pruning for Deep Neural Networks
- Activate Me!: Designing Efficient Activation Functions for Privacy-Preserving Machine Learning with Fully Homomorphic Encryption
- Hierarchical Graph Feature Enhancement with Adaptive Frequency Modulation for Visual Recognition
- Learning Layout and Style Reconfigurable GANs for Controllable Image Synthesis
- Learning to Predict Trustworthiness with Steep Slope Loss
- CoMoNM: A Cost Modeling Framework for Compute-Near-Memory Systems
- Deciphering the 2016 U.S. Presidential Campaign in the Twitter Sphere: A Comparison of the Trumpists and Clintonists
- Bandicoot: A Templated C++ Library for GPU Linear Algebra
- NeMo: A Neuron-Level Modularizing-While-Training Approach for Decomposing DNN Models
- ML-Doctor: Holistic Risk Assessment of Inference Attacks Against Machine Learning Models
- Unifying Scale-Aware Depth Prediction and Perceptual Priors for Monocular Endoscope Pose Estimation and Tissue Reconstruction
- FedNS: Improving Federated Learning for collaborative image classification on mobile clients
- Master Thesis: Neural Sign Language Translation by Learning Tokenization
- Two-Level Residual Distillation based Triple Network for Incremental Object Detection
- Zero-Shot Recognition through Image-Guided Semantic Classification
- Multi-GPU Training of ConvNets
- Hybrid-Hierarchical Fashion Graph Attention Network for Compatibility-Oriented and Personalized Outfit Recommendation
- health effects of dielectric gases: preliminary report
- Understanding the Error in Evaluating Adversarial Robustness
- Mobile-Friendly Deep Learning for Plant Disease Detection: A Lightweight CNN Benchmark Across 101 Classes of 33 Crops
- Self-Paced Uncertainty Estimation for One-shot Person Re-Identification
- NASA: Neural Articulated Shape Approximation
- Dissecting Generalized Category Discovery: Multiplex Consensus under Self-Deconstruction
- Novel View Synthesis using DDIM Inversion
- Predicting the Mumble of Wireless Channel with Sequence-to-Sequence Models
- EventNet: Asynchronous Recursive Event Processing
- A novel data-driven approach for transient stability prediction of power systems considering the operational variability
- ModiPick: SLA-aware Accuracy Optimization For Mobile Deep Inference
- HyperTea: A Hypergraph-based Temporal Enhancement and Alignment Network for Moving Infrared Small Target Detection
- Processing and acquisition traces in visual encoders: What does CLIP know about your camera?
- CRISP: Contrastive Residual Injection and Semantic Prompting for Continual Video Instance Segmentation
- Rethink ReLU to Training Better CNNs
- Large Model Empowered Embodied AI: A Survey on Decision-Making and Embodied Learning
- 3D latent diffusion models for parameterizing and history matching multiscenario facies systems
- From Pixel to Mask: A Survey of Out-of-Distribution Segmentation
- SynBrain: Enhancing Visual-to-fMRI Synthesis via Probabilistic Representation Learning
- Deep Sparse Subspace Clustering
- Non-uniqueness phenomenon of object representation in modelling IT cortex by deep convolutional neural network (DCNN)
- Convolutional Hough Matching Networks for Robust and Efficient Visual Correspondence
- Improving Automated COVID-19 Grading with Convolutional Neural Networks in Computed Tomography Scans: An Ablation Study
- Improving Robustness and Generality of NLP Models Using Disentangled Representations
- A System-Level Solution for Low-Power Object Detection
- Analyzing and Mitigating the Impact of Permanent Faults on a Systolic Array Based Neural Network Accelerator
- Structured Label Inference for Visual Understanding
- Joint Object and Part Segmentation using Deep Learned Potentials
- Hardware-Efficient Structure of the Accelerating Module for Implementation of Convolutional Neural Network Basic Operation
- Explainable AI Technique in Lung Cancer Detection Using Convolutional Neural Networks
- ALTIS: Modernizing GPGPU Benchmarking
- Lifted Relational Neural Networks
- TextureWGAN: Texture Preserving WGAN with MLE Regularizer for Inverse Problems
- Adaptively Denoising Proposal Collection for Weakly Supervised Object Localization
- IPG: Incremental Patch Generation for Generalized Adversarial Patch Training
- Hey Human, If your Facial Emotions are Uncertain, You Should Use Bayesian Neural Networks!
- ICD Coding from Clinical Text Using Multi-Filter Residual Convolutional Neural Network
- Temporal Proximity induces Attributes Similarity
- NIRMAL Pooling: An Adaptive Max Pooling Approach with Non-linear Activation for Enhanced Image Classification
- MixSearch: Searching for Domain Generalized Medical Image Segmentation Architectures
- Robust Pollen Imagery Classification with Generative Modeling and Mixup Training
- Rip van Winkle's Razor: A Simple Estimate of Overfit to Test Data
- Spurious Local Minima Are Common for Deep Neural Networks with Piecewise Linear Activations
- Drop-Activation: Implicit Parameter Reduction and Harmonic Regularization
- Source Printer Identification from Document Images Acquired using Smartphone
- How benign is benign overfitting?
- DeepFeatIoT: Unifying Deep Learned, Randomized, and LLM Features for Enhanced IoT Time Series Sensor Data Classification in Smart Industries
- Fast Haar Transforms for Graph Neural Networks
- Modeling the Uncertainty in Electronic Health Records: a Bayesian Deep Learning Approach
- Wide Neural Networks Forget Less Catastrophically
- MPT: Motion Prompt Tuning for Micro-Expression Recognition
- Deep Learning for Automated Identification of Vietnamese Timber Species: A Tool for Ecological Monitoring and Conservation
- Towards Train-Test Consistency for Semi-supervised Temporal Action Localization
- A Finer Calibration Analysis for Adversarial Robustness
- Large Language Models Show Signs of Alignment with Human Neurocognition During Abstract Reasoning
- NSGA-Net: Neural Architecture Search using Multi-Objective Genetic Algorithm
- Generating Training Data for Denoising Real RGB Images via Camera Pipeline Simulation
- Person Identification with Visual Summary for a Safe Access to a Smart Home
- Non-Parametric Calibration for Classification
- PatchGame: Learning to Signal Mid-level Patches in Referential Games
- Efficient motion-based metrics for video frame interpolation
- UniConvNet: Expanding Effective Receptive Field while Maintaining Asymptotically Gaussian Distribution for ConvNets of Any Scale
- Geometry-Aware Global Feature Aggregation for Real-Time Indirect Illumination
- Continual Learning in Neural Networks
- ShuffleNet V2: Practical Guidelines for Efficient CNN Architecture Design
- Deep Transform: Cocktail Party Source Separation via Complex Convolution in a Deep Neural Network
- Machine Learning Applications for Precision Agriculture: A Comprehensive Review
- Peer-Assisted Robotic Learning: A Data-Driven Collaborative Learning Approach for Cloud Robotic Systems
- Collegial Ensembles
- Transfer Learning for Melanoma Detection: Participation in ISIC 2017 Skin Lesion Classification Challenge
- Toward Lifelong Learning in Equilibrium Propagation: Sleep-like and Awake Rehearsal for Enhanced Stability
- Hashing as Tie-Aware Learning to Rank
- SIXray : A Large-scale Security Inspection X-ray Benchmark for Prohibited Item Discovery in Overlapping Images
- Learning a Unified Embedding for Visual Search at Pinterest
- AI-Skin : Skin Disease Recognition based on Self-learning and Wide Data Collection through a Closed Loop Framework
- PAC-GAN: An Effective Pose Augmentation Scheme for Unsupervised Cross-View Person Re-identification
- Image selective encryption analysis using mutual information in CNN based embedding space
- Forecasting Transportation Network Speed Using Deep Capsule Networks with Nested LSTM Models
- Continuous Perception for Classifying Shapes and Weights of Garmentsfor Robotic Vision Applications
- Automated Decision-based Adversarial Attacks
- Controlling Covariate Shift using Balanced Normalization of Weights
- Flexible Dataset Distillation: Learn Labels Instead of Images
- Implementation of Deep Neural Networks to Classify EEG Signals using Gramian Angular Summation Field for Epilepsy Diagnosis
- Sparse tree-based initialization for neural networks
- Reinforcement learning for batch bioprocess optimization
- Improving bug localization with word embedding and enhanced convolutional neural networks
- Defending Against Image Corruptions Through Adversarial Augmentations
- Interpretable CNNs for Object Classification
- Regularized Adaptation for Stable and Efficient Continuous-Level Learning on Image Processing Networks
- VulTriNet: A software vulnerability detection method based on tri-channel network
- Bidirectional LSTM-CRF for Clinical Concept Extraction
- An Intelligent Group Event Recommendation System in Social networks
- Understanding the Limitations of Variational Mutual Information Estimators
- CNN Acceleration by Low-rank Approximation with Quantized Factors
- Detecting Mislabeled and Corrupted Data via Pointwise Mutual Information
- Self-Supervised Representation Learning for Visual Anomaly Detection
- Active Learning for Breast Cancer Identification
- Cooperative Bi-path Metric for Few-shot Learning
- Comparison-Based Convolutional Neural Networks for Cervical Cell/Clumps Detection in the Limited Data Scenario
- Analysis of Video Feature Learning in Two-Stream CNNs on the Example of Zebrafish Swim Bout Classification
- Lightweight Multi-Scale Feature Extraction with Fully Connected LMF Layer for Salient Object Detection
- Why Does Stochastic Gradient Descent Slow Down in Low-Precision Training?
- A Bayesian Data Augmentation Approach for Learning Deep Models
- Multi-scale Domain-adversarial Multiple-instance CNN for Cancer Subtype Classification with Unannotated Histopathological Images
- PU-Net: Point Cloud Upsampling Network
- Spatial-Separated Curve Rendering Network for Efficient and High-Resolution Image Harmonization
- Learning to Generate 3D Shapes with Generative Cellular Automata
- Continuous Transition: Improving Sample Efficiency for Continuous Control Problems via MixUp
- Denoising and Verification Cross-Layer Ensemble Against Black-box Adversarial Attacks
- MISS: Multi-Interest Self-Supervised Learning Framework for Click-Through Rate Prediction
- Diffeomorphic Neural Operator Learning
- Confounder Identification-free Causal Visual Feature Learning
- Position: Ideas Should be the Center of Machine Learning Research
- Multimodal learning with next-token prediction for large multimodal models
- Deep Convolutional Decision Jungle for Image Classification
- Lung sounds classification using convolutional neural networks
- Application of Deep Learning in Fundus Image Processing for Ophthalmic Diagnosis -- A Review
- Medical knowledge embedding based on recursive neural network for multi-disease diagnosis
- Enhancing Recognition and Categorization of Skin Lesions with Tailored Deep Convolutional Networks and Robust Data Augmentation Techniques
- LU-Net: a multi-task network to improve the robustness of segmentation of left ventriclular structures by deep learning in 2D echocardiography
- LSM: Learning Subspace Minimization for Low-level Vision
- Divergent Search for Few-Shot Image Classification
- Fast Calculation of Probabilistic Power Flow: A Model-based Deep Learning Approach
- CornerNet: Detecting Objects as Paired Keypoints
- Efficient Cloth Simulation using Miniature Cloth and Upscaling Deep Neural Networks
- Reinforced Evolutionary Neural Architecture Search
- Alleviating Mode Collapse in GAN via Diversity Penalty Module
- Deep joint demosaicking and denoising
- Affective Image Content Analysis: Two Decades Review and New Perspectives
- SiamRPN++: Evolution of Siamese Visual Tracking with Very Deep Networks
- Fully Automatic Wound Segmentation with Deep Convolutional Neural Networks
- Parallel Deep Neural Networks Have Zero Duality Gap
- Facial Information Analysis Technology for Gender and Age Estimation
- Deep learning applications advance plant genomics research
- Being-ahead: Benchmarking and Exploring Accelerators for Hardware-Efficient AI Deployment
- Learning Representations of Satellite Images with Evaluations on Synoptic Weather Events
- Recurrent Deep Differentiable Logic Gate Networks
- Zero-Shot Learning by Convex Combination of Semantic Embeddings
- Adaptive Verifiable Training Using Pairwise Class Similarity
- Synthetic Data Generation for Emotional Depth Faces: Optimizing Conditional DCGANs via Genetic Algorithms in the Latent Space and Stabilizing Training with Knowledge Distillation
- Multi-view Gaze Target Estimation
- Re-evaluating Evaluation
- Cognitive computational neuroscience
- Regularized Evolutionary Population-Based Training
- Optimising the Performance of Convolutional Neural Networks across Computing Systems using Transfer Learning
- A geometry-inspired decision-based attack
- Keep It Real: Challenges in Attacking Compression-Based Adversarial Purification
- Input Invex Neural Network
- A Study of Gender Classification Techniques Based on Iris Images: A Deep Survey and Analysis
- Real-time Detection of Practical Universal Adversarial Perturbations
- Multi-Labelled Value Networks for Computer Go
- Rotation Equivariant Arbitrary-scale Image Super-Resolution
- Digital Twin Channel-Aided CSI Prediction: An Environment-Based Subspace Extraction Approach for Achieving Low Overhead and High Robustness
- ULU: A Unified Activation Function
- Towards Imperceptible Universal Attacks on Texture Recognition
- Tesserae: Scalable Placement Policies for Deep Learning Workloads
- Self-Error Adjustment: Theory and Practice of Balancing Individual Performance and Diversity in Ensemble Learning
- Smaller Models, Better Generalization
- Toward Errorless Training ImageNet-1k
- Pulmonary embolism identification in computerized tomography pulmonary angiography scans with deep learning technologies in COVID-19 patients
- Σ-net: Systematic Evaluation of Iterative Deep Neural Networks for Fast Parallel MR Image Reconstruction
- SoildNet: Soiling Degradation Detection in Autonomous Driving
- Combined Image Data Augmentations diminish the benefits of Adaptive Label Smoothing
- Robust Processing-In-Memory Neural Networks via Noise-Aware Normalization
- Metric Learning in an RKHS
- PoreFlow-Net: A 3D convolutional neural network to predict fluid flow through porous media
- DeePore: a deep learning workflow for rapid and comprehensive characterization of porous materials
- Adaptive Neuron-wise Discriminant Criterion and Adaptive Center Loss at Hidden Layer for Deep Convolutional Neural Network
- From Flat to Round: Redefining Brain Decoding with Surface-Based fMRI and Cortex Structure
- A2Mamba: Attention-augmented State Space Models for Visual Recognition
- Automated ultrasound doppler angle estimation using deep learning
- Energy-Efficient Real-Time 4-Stage Sleep Classification at 10-Second Resolution: A Comprehensive Study
- Age-Diverse Deepfake Dataset: Bridging the Age Gap in Deepfake Detection
- Slice or the Whole Pie? Utility Control for AI Models
- Scaling-Translation-Equivariant Networks with Decomposed Convolutional Filters
- Ethics through the Facets of Artificial Intelligence
- Deep Cross Residual Learning for Multitask Visual Recognition
- Learning to Combat Noisy Labels via Classification Margins
- Deep learning framework for crater detection and identification on the Moon and Mars
- Training Deep Neural Networks via Branch-and-Bound
- A Camera free fiber speckle wavemeter
- Robust Morph-Detection at Automated Border Control Gate using Deep Decomposed 3D Shape and Diffuse Reflectance
- Automatic Detection of Coronavirus Disease (COVID-19) in X-ray and CT Images: A Machine Learning-Based Approach
- CADD: Context aware disease deviations via restoration of brain images using normative conditional diffusion models
- Relationship-Embedded Representation Learning for Grounding Referring Expressions
- Deep Reasoning with Multi-Scale Context for Salient Object Detection
- A Scalable Optimization Mechanism for Pairwise based Discrete Hashing
- Zero Shot Domain Adaptive Semantic Segmentation by Synthetic Data Generation and Progressive Adaptation
- The Power of Many: Synergistic Unification of Diverse Augmentations for Efficient Adversarial Robustness
- QuantNet: Learning to Quantize by Learning within Fully Differentiable Framework
- RoIFusion: 3D Object Detection from LiDAR and Vision
- Where and How to Enhance: Discovering Bit-Width Contribution for Mixed Precision Quantization
- TrackNet: A Deep Learning Network for Tracking High-speed and Tiny Objects in Sports Applications
- The Geometry of Cortical Computation: Manifold Disentanglement and Predictive Dynamics in VCNet
- Fully-Convolutional Siamese Networks for Object Tracking
- Domain Adaptation with Incomplete Target Domains
- Toward Practical Equilibrium Propagation: Brain-inspired Recurrent Neural Network with Feedback Regulation and Residual Connections
- Coronary Artery Segmentation in Angiographic Videos Using A 3D-2D CE-Net
- Infrared Object Detection with Ultra Small ConvNets: Is ImageNet Pretraining Still Useful?
- Evaluation and Analysis of Deep Neural Transformers and Convolutional Neural Networks on Modern Remote Sensing Datasets
- Graph Classification Based on Skeleton and Component Features
- Towards Compact and Robust Deep Neural Networks
- FAIR-Pruner: Leveraging Tolerance of Difference for Flexible Automatic Layer-Wise Neural Network Pruning
- After the Party: Navigating the Mapping From Color to Ambient Lighting
- Deep Learning Face Representation by Joint Identification-Verification
- Fast Neural Network Adaptation via Parameter Remapping and Architecture Search
- Tackling Ill-posedness of Reversible Image Conversion with Well-posed Invertible Network
- Combining Ensembles and Data Augmentation can Harm your Calibration
- Defending Against Adversarial Attacks Using Random Forests
- GlaBoost: A multimodal Structured Framework for Glaucoma Risk Stratification
- Composite Quantization
- Improving Noise Efficiency in Privacy-preserving Dataset Distillation
- RaftMLP: How Much Can Be Done Without Attention and with Less Spatial Locality?
- Pulse Shape Discrimination Algorithms: Survey and Benchmark
- Discrete Rotation Equivariance for Point Cloud Recognition
- Set Pivot Learning: Redefining Generalized Segmentation with Vision Foundation Models
- Deep Built-Structure Counting in Satellite Imagery Using Attention Based Re-Weighting
- Reconstructing Trust Embeddings from Siamese Trust Scores: A Direct-Sum Approach with Fixed-Point Semantics
- Intensity augmentation for domain transfer of whole breast segmentation in MRI
- Increasing the Generalisation Capacity of Conditional VAEs
- NemaNet: A convolutional neural network model for identification of nematodes soybean crop in brazil
- Robust Visual Knowledge Transfer via EDA
- Translating Math Formula Images to LaTeX Sequences Using Deep Neural Networks with Sequence-level Training
- Partial observations and conservation laws: Grey-box modeling in biotechnology and optogenetics
- Improving 3D Object Detection through Progressive Population Based Augmentation
- Yelp Food Identification via Image Feature Extraction and Classification
- Fusion Sampling Validation in Data Partitioning for Machine Learning
- A Simple and Effective Method for Uncertainty Quantification and OOD Detection
- Imbalanced Deep Learning by Minority Class Incremental Rectification
- Residual-CNDS for Grand Challenge Scene Dataset
- Entropy-Based Uncertainty Calibration for Generalized Zero-Shot Learning
- Learning Competitive and Discriminative Reconstructions for Anomaly Detection
- Towards Higher Effective Rank in Parameter-efficient Fine-tuning using Khatri--Rao Product
- Initialization Strategies of Spatio-Temporal Convolutional Neural Networks
- Dream, Lift, Animate: From Single Images to Animatable Gaussian Avatars
- UIS-Mamba: Exploring Mamba for Underwater Instance Segmentation via Dynamic Tree Scan and Hidden State Weaken
- Variational Autoencoder-Based Black-Box Adversarial Attack on Collaborative DNN Inference
- Sinusoidal Approximation Theorem for Kolmogorov-Arnold Networks
- ZipNet-GAN: Inferring Fine-grained Mobile Traffic Patterns via a Generative Adversarial Neural Network
- EMA Without the Lag: Bias-Corrected Iterate Averaging Schemes
- L-GTA: Latent Generative Modeling for Time Series Augmentation
- Beyond Linear Bottlenecks: Spline-Based Knowledge Distillation for Culturally Diverse Art Style Classification
- Learn molecular representations from large-scale unlabeled molecules for drug discovery
- Particle reconstruction of volumetric particle image velocimetry with strategy of machine learning
- Flow Contrastive Estimation of Energy-Based Models
- Fractional Skipping: Towards Finer-Grained Dynamic CNN Inference
- Do We Need Fully Connected Output Layers in Convolutional Networks?
- Foundations and Models in Modern Computer Vision: Key Building Blocks in Landmark Architectures
- Initialization Using Perlin Noise for Training Networks with a Limited Amount of Data
- Your Spending Needs Attention: Modeling Financial Habits with Transformers
- Efficient and Robust Parallel DNN Training through Model Parallelism on Multi-GPU Platform
- How deep should be the depth of convolutional neural networks: a backyard dog case study
- CADDA: Class-wise Automatic Differentiable Data Augmentation for EEG Signals
- SemiNLL: A Framework of Noisy-Label Learning by Semi-Supervised Learning
- Describing Unseen Videos via Multi-Modal Cooperative Dialog Agents
- Graph Lineages and Skeletal Graph Products
- An Artificial Intelligence-Based System to Assess Nutrient Intake for Hospitalised Patients
- JQF: Optimal JPEG Quantization Table Fusion by Simulated Annealing on Texture Images and Predicting Textures
- Investigating the Invertibility of Multimodal Latent Spaces: Limitations of Optimization-Based Methods
- Tricks and Plug-ins for Gradient Boosting in Image Classification
- Multi-lane Detection Using Instance Segmentation and Attentive Voting
- Robust Collaborative Learning of Patch-level and Image-level Annotations for Diabetic Retinopathy Grading from Fundus Image
- Segment Anything for Video: A Comprehensive Review of Video Object Segmentation and Tracking from Past to Future
- Machine Learning for Detecting Data Exfiltration: A Review
- Explaining Natural Language Processing Classifiers with Occlusion and Language Modeling
- A Survey on Efficiency Optimization Techniques for DNN-based Video Analytics: Process Systems, Algorithms, and Applications
- DACA-Net: A Degradation-Aware Conditional Diffusion Network for Underwater Image Enhancement
- Crowding in humans is unlike that in convolutional neural networks
- You Only Look at One Sequence: Rethinking Transformer in Vision through Object Detection
- Agentic Privacy-Preserving Machine Learning
- RCR-AF: Enhancing Model Generalization via Rademacher Complexity Reduction Activation Function
- Training a Binary Weight Object Detector by Knowledge Transfer for Autonomous Driving
- Knowledge-augmented Column Networks: Guiding Deep Learning with Advice
- StarNet: Targeted Computation for Object Detection in Point Clouds
- FDNAS: Improving Data Privacy and Model Diversity in AutoML
- Synthesizing Optimal Parallelism Placement and Reduction Strategies on Hierarchical Systems for Deep Learning
- TIR-Diffusion: Diffusion-based Thermal Infrared Image Denoising via Latent and Wavelet Domain Optimization
- Moiré Zero: An Efficient and High-Performance Neural Architecture for Moiré Removal
- Object Recognition Datasets and Challenges: A Review
- From Waveforms to Pixels: A Survey on Audio-Visual Segmentation
- CTG-Insight: A Multi-Agent Interpretable LLM Framework for Cardiotocography Analysis and Classification
- AI in Agriculture: A Survey of Deep Learning Techniques for Crops, Fisheries and Livestock
- Shallow Deep Learning Can Still Excel in Fine-Grained Few-Shot Learning
- SwinECAT: A Transformer-based fundus disease classification model with Shifted Window Attention and Efficient Channel Attention
- Model compression as constrained optimization, with application to neural nets. Part V: combining compressions
- Noisy Gradient Descent Converges to Flat Minima for Nonconvex Matrix Factorization
- Automatic Building Extraction in Aerial Scenes Using Convolutional Networks
- NUMA-aware FFT-based Convolution on ARMv8 Many-core CPUs
- Deep Representation of Facial Geometric and Photometric Attributes for Automatic 3D Facial Expression Recognition
- Zero-Shot Machine Unlearning with Proxy Adversarial Data Generation
- Efficient Neural Architecture Transformation Searchin Channel-Level for Object Detection
- Exploring the Link Between Bayesian Inference and Embodied Intelligence: Toward Open Physical-World Embodied AI Systems
- Boost Self-Supervised Dataset Distillation via Parameterization, Predefined Augmentation, and Approximation
- Shaping the future of myopia with artificial intelligence: Mapping trends and promising directions
- A Frank-Wolfe Framework for Efficient and Effective Adversarial Attacks
- Quantum optical shallow networks
- MGN-Net: a multi-view graph normalizer for integrating heterogeneous biological network populations
- Mask-Free Audio-driven Talking Face Generation for Enhanced Visual Quality and Identity Preservation
- Transferability of Adversarial Examples to Attack Cloud-based Image Classifier Service
- Demystifying Learning Rate Policies for High Accuracy Training of Deep Neural Networks
- An Adaptive Random Fourier Features approach Applied to Learning Stochastic Differential Equations
- Embracing the Unreliability of Memory Devices for Neuromorphic Computing
- Progressive Cluster Purification for Transductive Few-shot Learning
- Class Subset Selection for Transfer Learning using Submodularity
- Deep learning based mood tagging for Chinese song lyrics
- Learning to Read by Spelling: Towards Unsupervised Text Recognition
- Adversarial Structure Matching for Structured Prediction Tasks
- Fine-grained Recognition Datasets for Biodiversity Analysis
- Aggregating Deep Convolutional Features for Image Retrieval
- Benefits of Feature Extraction and Temporal Sequence Analysis for Video Frame Prediction: An Evaluation of Hybrid Deep Learning Models
- Unsupervised Visual Representation Learning by Online Constrained K-Means
- Adaptive Fuzzy Time Series Forecasting via Partially Asymmetric Convolution and Sub-Sliding Window Fusion
- Modelling Diffuse Subcellular Protein Structures as Dynamic Social Networks
- Denoising convolutional autoencoder based B-mode ultrasound tongue image feature extraction
- FED-PsyAU: Privacy-Preserving Micro-Expression Recognition via Psychological AU Coordination and Dynamic Facial Motion Modeling
- Telugu OCR Framework using Deep Learning
- Beyond Class Tokens: LLM-guided Dominant Property Mining for Few-shot Classification
- Your Attention Matters: to Improve Model Robustness to Noise and Spurious Correlations
- Surrogate-assisted Particle Swarm Optimisation for Evolving Variable-length Transferable Blocks for Image Classification
- LTD: Low Temperature Distillation for Gradient Masking-free Adversarial Training
- Learning to Branch for Multi-Task Learning
- Advances and Challenges in Deep Lip Reading
- A Sparse Coding Interpretation of Neural Networks and Theoretical Implications
- TDAPNet: Prototype Network with Recurrent Top-Down Attention for Robust Object Classification under Partial Occlusion
- Generative Pre-training for Subjective Tasks: A Diffusion Transformer-Based Framework for Facial Beauty Prediction
- Computational Advantages of Multi-Grade Deep Learning: Convergence Analysis and Performance Insights
- Deep Neural Networks for Blind Image Quality Assessment: Addressing the Data Challenge
- Learning Smooth Representation for Unsupervised Domain Adaptation
- Real-time PCG Anomaly Detection by Adaptive 1D Convolutional Neural Networks
- CB2CF: A Neural Multiview Content-to-Collaborative Filtering Model for Completely Cold Item Recommendations
- Giving Up Control: Neurons as Reinforcement Learning Agents
- Robust Deep Graph Based Learning for Binary Classification
- Neural Networks Are More Productive Teachers Than Human Raters: Active Mixup for Data-Efficient Knowledge Distillation from a Blackbox Model
- MiLeNAS: Efficient Neural Architecture Search via Mixed-Level Reformulation
- SuperNet -- An efficient method of neural networks ensembling
- ShrinkTeaNet: Million-scale Lightweight Face Recognition via Shrinking Teacher-Student Networks
- SaltiNet: Scan-path Prediction on 360 Degree Images using Saliency Volumes
- Color histogram equalization and fine-tuning to improve expression recognition of (partially occluded) faces on sign language datasets
- Fast Crack Detection Using Convolutional Neural Network
- Machine Learning With Neuromorphic Photonics
- Dataset Distillation with Infinitely Wide Convolutional Networks
- A roadmap for AI in robotics
- A mini-batch training strategy for deep subspace clustering networks
- Rethinking on Multi-Stage Networks for Human Pose Estimation
- StressNet: Deep Learning to Predict Stress With Fracture Propagation in Brittle Materials
- Compact Hash Code Learning with Binary Deep Neural Network
- Visual Tracking via Dynamic Memory Networks
- Do We Need Sound for Sound Source Localization?
- Towards Understanding the Regularization of Adversarial Robustness on Neural Networks
- Data Augmentation for Enhancing EEG-based Emotion Recognition with Deep Generative Models
- A Systematic Literature Review on the Use of Deep Learning in Software Engineering Research
- Learning Spatio-Temporal Features with Two-Stream Deep 3D CNNs for Lipreading
- Adaptive Label Smoothing
- Role of Intonation in Scoring Spoken English
- FBNetV3: Joint Architecture-Recipe Search using Predictor Pretraining
- Probabilistic bounds on neuron death in deep rectifier networks
- Intra-Model Collaborative Learning of Neural Networks
- Putting visual object recognition in context
- Demographic-aware fine-grained visual recognition of pediatric wrist pathologies
- Quadratic Suffices for Over-parametrization via Matrix Chernoff Bound
- SegFormer: Simple and Efficient Design for Semantic Segmentation with Transformers
- Higher-order Graph Convolutional Networks
- Scale-aware Automatic Augmentation for Object Detection
- TextTopicNet - Self-Supervised Learning of Visual Features Through Embedding Images on Semantic Text Spaces
- YNU-HPCC at SemEval-2020 Task 8: Using a Parallel-Channel Model for Memotion Analysis
- Deep learning for lithological classification of carbonate rock micro-CT images
- Localizing Discriminative Visual Landmarks for Place Recognition
- Neural-Symbolic Descriptive Action Model from Images: The Search for STRIPS
- BRIEF: Backward Reduction of CNNs with Information Flow Analysis
- Neural Graph Embedding Methods for Natural Language Processing
- On Generalization Bounds of a Family of Recurrent Neural Networks
- Edge-Semantic Learning Strategy for Layout Estimation in Indoor Environment
- Non-Gaussianity of Stochastic Gradient Noise
- Efficient Structured Pruning and Architecture Searching for Group Convolution
- Deep learning for Chemometric and non-translational data
- Norm-preserving Orthogonal Permutation Linear Unit Activation Functions (OPLU)
- Towards a Robust Deep Neural Network in Texts: A Survey
- Universal approximations of permutation invariant/equivariant functions by deep neural networks
- Learning by Design: Structuring and Documenting the Human Choices in Machine Learning Development
- Virne: A Comprehensive Benchmark for RL-based Network Resource Allocation in NFV
- Semantic Image Segmentation via Deep Parsing Network
- DeepSentiBank: Visual Sentiment Concept Classification with Deep Convolutional Neural Networks
- POIRot: A rotation invariant omni-directional pointnet
- MC-SSL0.0: Towards Multi-Concept Self-Supervised Learning
- YOLO for Knowledge Extraction from Vehicle Images: A Baseline Study
- Large-time asymptotics in deep learning
- Handling Out-of-Distribution Data: A Survey
- Target Driven Instance Detection
- Low-memory convolutional neural networks through incremental depth-first processing
- Multi-Bias Non-linear Activation in Deep Neural Networks
- Detecting Driveable Area for Autonomous Vehicles
- Improving the HardNet Descriptor
- On Arbitrary Predictions from Equally Valid Models
- RealDeal: Enhancing Realism and Details in Brain Image Generation via Image-to-Image Diffusion Models
- Group Activity Prediction with Sequential Relational Anticipation Model
- Class-dependent Compression of Deep Neural Networks
- A Closed-Form Learned Pooling for Deep Classification Networks
- RECAL: Reuse of Established CNN classifer Apropos unsupervised Learning paradigm
- The synergy of neuromarketing and artificial intelligence: A comprehensive literature review in the last decade
- The Dynamics of Handwriting Improves the Automated Diagnosis of Dysgraphia
- Mixture separability loss in a deep convolutional network for image classification
- Collaborative Spatio-temporal Feature Learning for Video Action Recognition
- BiPointNet: Binary Neural Network for Point Clouds
- Point Adversarial Self Mining: A Simple Method for Facial Expression Recognition
- SemiSegECG: A Multi-Dataset Benchmark for Semi-Supervised Semantic Segmentation in ECG Delineation
- Backdoor Embedding in Convolutional Neural Network Models via Invisible Perturbation
- ChronoSelect: Robust Learning with Noisy Labels via Dynamics Temporal Memory
- Multi-adversarial Faster-RCNN for Unrestricted Object Detection
- Trade-off between deep learning for species identification and inference about predator-prey co-occurrence
- Scaling Recommender Transformers to One Billion Parameters
- On Learning Where To Look
- Iwin Transformer: Hierarchical Vision Transformer using Interleaved Windows
- Convolutional Nonlinear Dictionary with Cascaded Structure Filter Banks
- Mid-level Elements for Object Detection
- Patchy Image Structure Classification Using Multi-Orientation Region Transform
- Handcrafted vs Deep Learning Classification for Scalable Video QoE Modeling
- The Impact of GPU DVFS on the Energy and Performance of Deep Learning: an Empirical Study
- Deep Representation Learning in Speech Processing: Challenges, Recent Advances, and Future Trends
- A Semantic Indexing Structure for Image Retrieval
- Window-Object Relationship Guided Representation Learning for Generic Object Detections
- Visual Place Recognition for Large-Scale UAV Applications
- Object segmentation in the wild with foundation models: application to vision assisted neuro-prostheses for upper limbs
- Audio Visual Emotion Recognition with Temporal Alignment and Perception Attention
- Towards calibrated and scalable uncertainty representations for neural networks
- AIDA: Associative DNN Inference Accelerator
- Memory Matters: Convolutional Recurrent Neural Network for Scene Text Recognition
- Quantum Machine Learning Playground
- Dissecting the effectiveness of deep features as metric of perceptual image quality
- Why Divisive Normalization works in image segmentation?
- FINN-R: An End-to-End Deep-Learning Framework for Fast Exploration of Quantized Neural Networks
- Traversing the noise of dynamic mini-batch sub-sampled loss functions: A visual guide
- Joint Spatial and Layer Attention for Convolutional Networks
- Reconstructing A Large Scale 3D Face Dataset for Deep 3D Face Identification
- Joint Learning of Distributed Representations for Images and Texts
- Domain2Vec: Domain Embedding for Unsupervised Domain Adaptation
- Federated Learning with Domain Generalization
- Applying deep learning to classify pornographic images and videos
- Opportunities for Analog Coding in Emerging Memory Systems
- Aspect-based Opinion Summarization with Convolutional Neural Networks
- Fine-Grained Classification via Mixture of Deep Convolutional Neural Networks
- Hierarchical Invariant Feature Learning with Marginalization for Person Re-Identification
- On Generalization Error Bounds of Noisy Gradient Methods for Non-Convex Learning
- DARI: Distance metric And Representation Integration for Person Verification
- Comparing Classes of Estimators: When does Gradient Descent Beat Ridge Regression in Linear Models?
- Understanding Deep Image Representations by Inverting Them
- Situational Fusion of Visual Representation for Visual Navigation
- NEUZZ: Efficient Fuzzing with Neural Program Smoothing
- A Noise-Sensitivity-Analysis-Based Test Prioritization Technique for Deep Neural Networks
- Deep learning empowers genomic selection of pest-resistant grapevine
- Microscopic image processing platform for multi-class cell segmentation using deep learning
- Weakly Supervised PET Tumor Detection Using Class Response
- How Much Can We Really Trust You? Towards Simple, Interpretable Trust Quantification Metrics for Deep Neural Networks
- Deep multi-modal networks for book genre classification based on its cover
- Legal infrastructure for transformative AI governance
- Transfer Learning in Visual and Relational Reasoning
- A* Tree Search for Portfolio Management
- Multi-Language Identification Using Convolutional Recurrent Neural Network
- Parsimonious Bayesian deep networks
- Localization-aware Channel Pruning for Object Detection
- Training Meta-Surrogate Model for Transferable Adversarial Attack
- Unsupervised Person Re-identification via Softened Similarity Learning
- Learning Wake-Sleep Recurrent Attention Models
- Pinpointing the Memory Behaviors of DNN Training
- Deep Cross-Modal Hashing
- Machine learning in geo- and environmental sciences: From small to large scale
- Spatio-temporal Video Re-localization by Warp LSTM
- Clinical applications of deep learning in breast MRI
- A Novel Downsampling Strategy Based on Information Complementarity for Medical Image Segmentation
- Deep reinforcement learning for inventory control: A roadmap
- Dementia Severity Classification under Small Sample Size and Weak Supervision in Thick Slice MRI
- Exploring and Improving Mobile Level Vision Transformers
- PatchGuard: A Provably Robust Defense against Adversarial Patches via Small Receptive Fields and Masking
- Understanding Clipping for Federated Learning: Convergence and Client-Level Differential Privacy
- Evolutionary Ensemble Learning for Multivariate Time Series Prediction
- Data Sanity Check for Deep Learning Systems via Learnt Assertions
- Semantic Correlation Promoted Shape-Variant Context for Segmentation
- A Benchmark Comparison of Visual Place Recognition Techniques for Resource-Constrained Embedded Platforms
- Deep Learning with S-shaped Rectified Linear Activation Units
- Orthogonal Convolutional Neural Networks
- Creative Captioning: An AI Grand Challenge Based on the Dixit Board Game
- Efficient Hardware Realizations of Feedforward Artificial Neural Networks
- Boosting Black-Box Adversarial Attacks with Meta Learning
- Pathology-Aware Generative Adversarial Networks for Medical Image Augmentation
- Large Scale Font Independent Urdu Text Recognition System
- MetaConcept: Learn to Abstract via Concept Graph for Weakly-Supervised Few-Shot Learning
- Low-light Image Enhancement Algorithm Based on Retinex and Generative Adversarial Network
- CSGAN: Cyclic-Synthesized Generative Adversarial Networks for Image-to-Image Transformation
- Selecting Relevant Features from a Multi-domain Representation for Few-shot Classification
- How Much Does Audio Matter to Recognize Egocentric Object Interactions?
- On the Exploitation of Neuroevolutionary Information: Analyzing the Past for a More Efficient Future
- Locality Preserving Joint Transfer for Domain Adaptation
- Weakly Supervised Complementary Parts Models for Fine-Grained Image Classification from the Bottom Up
- AlexNet‐NDTL: Classification of MRI brain tumor images using modified AlexNet with deep transfer learning and Lipschitz‐based data augmentation
- Rectal cancer: Toward fully automatic discrimination of T2 and T3 rectal cancers using deep convolutional neural network
- Industry and Academic Research in Computer Vision
- Non-Local ConvLSTM for Video Compression Artifact Reduction
- Neural Architecture Dilation for Adversarial Robustness
- Feature Level Fusion from Facial Attributes for Face Recognition
- Contributions to Large Scale Bayesian Inference and Adversarial Machine Learning
- Matching Distributions via Optimal Transport for Semi-Supervised Learning
- Dynamic R-CNN: Towards High Quality Object Detection via Dynamic Training
- M-ar-K-Fast Independent Component Analysis
- Adaptive Consistency Regularization for Semi-Supervised Transfer Learning
- Graph-LDA: Graph Structure Priors to Improve the Accuracy in Few-Shot Classification
- Image Generators are Generalist Vision Learners
- Optimization and Generalization of Regularization-Based Continual Learning: a Loss Approximation Viewpoint
- Persistent Evidence of Local Image Properties in Generic ConvNets
- On Vectorization of Deep Convolutional Neural Networks for Vision Tasks
- Stingray Detection of Aerial Images Using Augmented Training Images Generated by A Conditional Generative Model
- Ensemble Learning based on Classifier Prediction Confidence and Comprehensive Learning Particle Swarm Optimisation for polyp localisation
- MetaAugment: Sample-Aware Data Augmentation Policy Learning
- Detection of Paroxysmal Atrial Fibrillation using Attention-based Bidirectional Recurrent Neural Networks
- VISALOGY: Answering Visual Analogy Questions
- Learning a Continuous Representation of 3D Molecular Structures with Deep Generative Models
- SurReal: Complex-Valued Learning as Principled Transformations on a Scaling and Rotation Manifold
- SiamVGG: Visual Tracking using Deeper Siamese Networks
- Improving Image Captioning by Leveraging Knowledge Graphs
- Detecting Adversarial Patches with Class Conditional Reconstruction Networks
- Pneumothorax Segmentation: Deep Learning Image Segmentation to predict Pneumothorax
- ACNe: Attentive Context Normalization for Robust Permutation-Equivariant Learning
- Online Algorithms and Policies Using Adaptive and Machine Learning Approaches
- Malware Detection Using Frequency Domain-Based Image Visualization and Deep Learning
- Learning from Extrinsic and Intrinsic Supervisions for Domain Generalization
- SafeAccess+: An Intelligent System to make Smart Home Safer and Americans with Disability Act Compliant
- Significant Wave Height Prediction based on Wavelet Graph Neural Network
- A GPU-Outperforming FPGA Accelerator Architecture for Binary Convolutional Neural Networks
- Weakly Aligned Cross-Modal Learning for Multispectral Pedestrian Detection
- What Can I Do Around Here? Deep Functional Scene Understanding for Cognitive Robots
- Altitude Training: Strong Bounds for Single-Layer Dropout
- Diabetic Retinopathy Detection via Deep Convolutional Networks for Discriminative Localization and Visual Explanation
- Tissue characterization based on the analysis on i3DUS data for diagnosis support in neurosurgery
- ORACLE: Optimized Radio clAssification through Convolutional neuraL nEtworks
- Point in, Box out: Beyond Counting Persons in Crowds
- Labeled Data Generation with Inexact Supervision
- Developing efficient transfer learning strategies for robust scene recognition in mobile robotics using pre-trained convolutional neural networks
- ToAlign: Task-oriented Alignment for Unsupervised Domain Adaptation
- Intelligent fault diagnosis of hydraulic piston pump combining improved LeNet-5 and PSO hyperparameter optimization
- Image Matters: Visually modeling user behaviors using Advanced Model Server
- Multi-task Supervised Learning via Cross-learning
- Local Model Feature Transformations
- Out-distribution aware Self-training in an Open World Setting
- Reappraising Domain Generalization in Neural Networks
- Adversarial Distillation for Ordered Top-k Attacks
- Interpretation of Feature Space using Multi-Channel Attentional Sub-Networks
- Passive Attention in Artificial Neural Networks Predicts Human Visual Selectivity
- Performance landscape of resource-constrained platforms targeting DNNs
- CelebHair: A New Large-Scale Dataset for Hairstyle Recommendation based on CelebA
- Spectral Tensor Train Parameterization of Deep Learning Layers
- Practical Relative Order Attack in Deep Ranking
- AutoEG: Automated Experience Grafting for Off-Policy Deep Reinforcement Learning
- Facilitating Connected Autonomous Vehicle Operations Using Space-weighted Information Fusion and Deep Reinforcement Learning Based Control
- The Costs and Benefits of Goal-Directed Attention in Deep Convolutional Neural Networks
- Large-scale Kernel Methods and Applications to Lifelong Robot Learning
- RTN: Reparameterized Ternary Network
- FOTS: Fast Oriented Text Spotting with a Unified Network
- Multilayer Collaborative Low-Rank Coding Network for Robust Deep Subspace Discovery
- ElixirNet: Relation-aware Network Architecture Adaptation for Medical Lesion Detection
- Modeling Image Virality with Pairwise Spatial Transformer Networks
- Sparse Transfer Learning via Winning Lottery Tickets
- Colored Kimia Path24 Dataset: Configurations and Benchmarks with Deep Embeddings
- Open-World Entity Segmentation
- BlackMarks: Blackbox Multibit Watermarking for Deep Neural Networks
- Negative Margin Matters: Understanding Margin in Few-shot Classification
- A multiscale neural network based on hierarchical matrices
- Image Augmentations for GAN Training
- Bridging the Gaps Between Residual Learning, Recurrent Neural Networks and Visual Cortex
- Places205-VGGNet Models for Scene Recognition
- Blind Adversarial Network Perturbations
- Electricity Theft Detection with self-attention
- Temporal Probability Calibration
- MDLdroid: a ChainSGD-reduce Approach to Mobile Deep Learning for Personal Mobile Sensing
- Residual Continual Learning
- MosaicOS: A Simple and Effective Use of Object-Centric Images for Long-Tailed Object Detection
- ChainerCV: a Library for Deep Learning in Computer Vision
- Measuring and Understanding Sensory Representations within Deep Networks Using a Numerical Optimization Framework
- Temperature check: theory and practice for training models with softmax-cross-entropy losses
- Splash: User-friendly Programming Interface for Parallelizing Stochastic Algorithms
- Faceness-Net: Face Detection through Deep Facial Part Responses
- Efficient Pre-trained Features and Recurrent Pseudo-Labeling in Unsupervised Domain Adaptation
- Lifelong Learning with Searchable Extension Units
- Self-Supervised Contextual Bandits in Computer Vision
- XAI-Guided Analysis of Residual Networks for Interpretable Pneumonia Detection in Paediatric Chest X-rays
- CALPA-NET: Channel-pruning-assisted Deep Residual Network for Steganalysis of Digital Images
- SapAugment: Learning A Sample Adaptive Policy for Data Augmentation
- In Proximity of ReLU DNN, PWA Function, and Explicit MPC
- AdaIN-Switchable CycleGAN for Efficient Unsupervised Low-Dose CT Denoising
- asya: Mindful verbal communication using deep learning
- RSINet: Rotation-Scale Invariant Network for Online Visual Tracking
- OpenEI: An Open Framework for Edge Intelligence
- Deep Heterogeneous Feature Fusion for Template-Based Face Recognition
- Muddling Label Regularization: Deep Learning for Tabular Datasets
- Yield Loss Reduction and Test of AI and Deep Learning Accelerators
- Contextual Prediction Difference Analysis for Explaining Individual Image Classifications
- Accelerating Very Deep Convolutional Networks for Classification and Detection
- In Defense of the Direct Perception of Affordances
- Quantum Boltzmann Machines using Parallel Annealing for Medical Image Classification
- The MVTec Anomaly Detection Dataset: A Comprehensive Real-World Dataset for Unsupervised Anomaly Detection
- Adversarial Learning with Margin-based Triplet Embedding Regularization
- Semantic and Visual Similarities for Efficient Knowledge Transfer in CNN Training
- Prediction of Carbon Nanostructure Mechanical Properties and Role of Defects Using Machine Learning
- Learning to Sort Image Sequences via Accumulated Temporal Differences
- Enhancing Multi-Robot Perception via Learned Data Association
- Implementation of Fruits Recognition Classifier using Convolutional Neural Network Algorithm for Observation of Accuracies for Various Hidden Layers
- Translational application of a self-organized deep feature engineering pipeline for non-invasive pulmonary hypertension classification from routine chest radiographs
- PASS: Peer-agreement based sample selection for training with instance dependent noisy labels
- SolarNet: A Deep Learning Framework to Map Solar Power Plants In China From Satellite Imagery
- Benchmarking Physical Performance of Neural Inference Circuits
- A Survey of Coded Distributed Computing
- Grasping Detection Network with Uncertainty Estimation for Confidence-Driven Semi-Supervised Domain Adaptation
- Depthwise Convolution is All You Need for Learning Multiple Visual Domains
- Optimizing the Whole-life Cost in End-to-end CNN Acceleration
- EKT: Exercise-Aware Knowledge Tracing for Student Performance Prediction
- Reducing Inference Latency with Concurrent Architectures for Image Recognition
- HiFT: Hierarchical Feature Transformer for Aerial Tracking
- Sionnx: Automatic Unit Test Generator for ONNX Conformance
- Generalized Dropout
- Efficient Convolutional Neural Network Training with Direct Feedback Alignment
- Ensembles of feedforward-designed convolutional neural networks
- Human Following for Wheeled Robot with Monocular Pan-tilt Camera
- GSPMD: General and Scalable Parallelization for ML Computation Graphs
- Image Aesthetics Assessment using Multi Channel Convolutional Neural Networks
- Deep Residual Learning in Spiking Neural Networks
- A deep learning model captures position-specific effects of plant regulatory sequences and suggests genes under complex regulation
- Ripple Attention for Visual Perception with Sub-quadratic Complexity
- Meta-Learning with Network Pruning
- OpEvo: An Evolutionary Method for Tensor Operator Optimization
- Large Learning Rates Simultaneously Achieve Robustness to Spurious Correlations and Compressibility
- Poisoned classifiers are not only backdoored, they are fundamentally broken
- Autoencoding Features for Aviation Machine Learning Problems
- Exponentially Increasing the Capacity-to-Computation Ratio for Conditional Computation in Deep Learning
- CNS-Bench: Benchmarking Image Classifier Robustness Under Continuous Nuisance Shifts
- Convolutional Neural Networks over Tree Structures for Programming Language Processing
- Evaluating Multimodal Representations on Visual Semantic Textual Similarity
- Joint 2D-3D Breast Cancer Classification
- Dataset Distillation as Data Compression: A Rate-Utility Perspective
- Triple Generative Adversarial Networks
- Synthesizing human-like sketches from natural images using a conditional convolutional decoder
- AttaNet: Attention-Augmented Network for Fast and Accurate Scene Parsing
- VGS-ATD: Robust Distributed Learning for Multi-Label Medical Image Classification Under Heterogeneous and Imbalanced Conditions
- DAMIA: Leveraging Domain Adaptation as a Defense against Membership Inference Attacks
- Learning from Scratch: Structurally-masked Transformer for Next Generation Lib-free Simulation
- Sharp Statistical Guarantees for Adversarially Robust Gaussian Classification
- Few-Shot Learning in Video and 3D Object Detection: A Survey
- CM-UNet: A Self-Supervised Learning-Based Model for Coronary Artery Segmentation in X-Ray Angiography
- Bringing Balance to Hand Shape Classification: Mitigating Data Imbalance Through Generative Models
- Fractional Spike Differential Equations Neural Network with Efficient Adjoint Parameters Training
- A Broader Study of Cross-Domain Few-Shot Learning
- Maximum Mutation Reinforcement Learning for Scalable Control
- Elastic Neural Networks for Classification
- X-volution: On the unification of convolution and self-attention
- Perception Matters: Exploring Imperceptible and Transferable Anti-forensics for GAN-generated Fake Face Imagery Detection
- Disguising Personal Identity Information in EEG Signals
- A PCA-Based Convolutional Network
- Features extraction for image identification using computer vision
- Cross-Modal Distillation For Widely Differing Modalities
- Adaptive Loss Function for Super Resolution Neural Networks Using Convex Optimization Techniques
- Multi-Agent Reinforcement Learning for Sample-Efficient Deep Neural Network Mapping
- Adaptive Relative Pose Estimation Framework with Dual Noise Tuning for Safe Approaching Maneuvers
- Scale Steerable Filters for Locally Scale-Invariant Convolutional Neural Networks
- GASPnet: Global Agreement to Synchronize Phases
- Deep Structured-Output Regression Learning for Computational Color Constancy
- Laparoscopy Surgery CO2 Removal via Generative Adversary Network and Dark Channel Prior
- Visual Language Modeling on CNN Image Representations
- Improve CAM with Auto-adapted Segmentation and Co-supervised Augmentation
- FishNet: A Camera Localizer using Deep Recurrent Networks
- Natural Language Guided Visual Relationship Detection
- Screen Gleaning: A Screen Reading TEMPEST Attack on Mobile Devices Exploiting an Electromagnetic Side Channel
- From SGD to Spectra: A Theory of Neural Network Weight Dynamics
- FedGA: A Fair Federated Learning Framework Based on the Gini Coefficient
- A Dictionary Approach to Domain-Invariant Learning in Deep Networks
- End-to-end training of deep kernel map networks for image classification
- Learning Context Graph for Person Search
- Impact of Deep Learning Assistance on the Histopathologic Review of Lymph Nodes for Metastatic Breast Cancer
- Classification Under Human Assistance
- Data-Free Knowledge Amalgamation via Group-Stack Dual-GAN
- HGC: Hierarchical Group Convolution for Highly Efficient Neural Network
- Variational Implicit Processes
- The Serial Scaling Hypothesis
- ReNet: A Recurrent Neural Network Based Alternative to Convolutional Networks
- Conditional diffusion models for guided anomaly detection in brain images using fluid-driven anomaly randomization
- Hybrid Ensemble Approaches: Optimal Deep Feature Fusion and Hyperparameter-Tuned Classifier Ensembling for Enhanced Brain Tumor Classification
- AD-GS: Object-Aware B-Spline Gaussian Splatting for Self-Supervised Autonomous Driving
- A Smoother Way to Train Structured Prediction Models
- Attributes Guided Feature Learning for Vehicle Re-identification
- Rethinking Spatial Dimensions of Vision Transformers
- Co-Design of Deep Neural Nets and Neural Net Accelerators for Embedded Vision Applications
- DUSE: A Data Expansion Framework for Low-resource Automatic Modulation Recognition based on Active Learning
- Adaptive Ultrasound Beamforming using Deep Learning
- A3GAN: An Attribute-aware Attentive Generative Adversarial Network for Face Aging
- Deep Relighting Networks for Image Light Source Manipulation
- Visual Security Evaluation of Learnable Image Encryption Methods against Ciphertext-only Attacks
- AttentionNAS: Spatiotemporal Attention Cell Search for Video Classification
- VideoClick: Video Object Segmentation with a Single Click
- Random Neural Networks in the Infinite Width Limit as Gaussian Processes
- A Survey on Human-aware Robot Navigation
- Optimization Methods for Large-Scale Machine Learning
- Online Training and Pruning of Deep Reinforcement Learning Networks
- High Performance and Portable Convolution Operators for ARM-based Multicore Processors
- Assaying Out-Of-Distribution Generalization in Transfer Learning
- Selective Embedding for Deep Learning
- Deep Metric Transfer for Label Propagation with Limited Annotated Data
- Hyperspectral Image Classification with Spatial Consistence Using Fully Convolutional Spatial Propagation Network
- Automated Scoring of Nuclear Pleomorphism Spectrum with Pathologist-level Performance in Breast Cancer
- Weakly-Supervised Monocular Depth Estimationwith Resolution-Mismatched Data
- Reinforcement Learning for Industrial Control Network Cyber Security Orchestration
- Auto-Compressing Networks
- Incentivised Orchestrated Training Architecture (IOTA): A Technical Primer for Release
- Speed/accuracy trade-offs for modern convolutional object detectors
- Adversarial Neural Pruning with Latent Vulnerability Suppression
- Learning to Sample the Most Useful Training Patches from Images
- Domain Adaptation by Maximizing Population Correlation with Neural Architecture Search
- Deep Hashing Learning for Visual and Semantic Retrieval of Remote Sensing Images
- TMA: Tera-MACs/W Neural Hardware Inference Accelerator with a Multiplier-less Massive Parallel Processor
- Ensemble Transfer Learning for Emergency Landing Field Identification on Moderate Resource Heterogeneous Kubernetes Cluster
- Inspect Transfer Learning Architecture with Dilated Convolution
- Deep Learning: Our Miraculous Year 1990-1991
- Transferring Inter-Class Correlation
- Understanding the Mechanism of Deep Learning Framework for Lesion Detection in Pathological Images with Breast Cancer
- Attention to Lesion: Lesion-Aware Convolutional Neural Network for Retinal Optical Coherence Tomography Image Classification
- Search to Distill: Pearls are Everywhere but not the Eyes
- HR-RCNN: Hierarchical Relational Reasoning for Object Detection
- On the Origin of Species of Self-Supervised Learning
- Revamping Cross-Modal Recipe Retrieval with Hierarchical Transformers and Self-supervised Learning
- F2GAN: Fusing-and-Filling GAN for Few-shot Image Generation
- ATSO: Asynchronous Teacher-Student Optimization for Semi-Supervised Medical Image Segmentation
- Improving Deep Neural Network with Multiple Parametric Exponential Linear Units
- Interpretable Complex-Valued Neural Networks for Privacy Protection
- Feature Combination Meets Attention: Baidu Soccer Embeddings and Transformer based Temporal Detection
- An Adversarial Learning Approach to Medical Image Synthesis for Lesion Detection
- Relaxed Conditional Image Transfer for Semi-supervised Domain Adaptation
- ObjectNet Dataset: Reanalysis and Correction
- TSInsight: A local-global attribution framework for interpretability in time-series data
- Deep Learning-Based Gait Recognition Using Smartphones in the Wild
- Exploiting Contextual Information with Deep Neural Networks
- Supervised Online Hashing via Hadamard Codebook Learning
- Local Label Propagation for Large-Scale Semi-Supervised Learning
- Launchpad: A Programming Model for Distributed Machine Learning Research
- NWT: Towards natural audio-to-video generation with representation learning
- Computer-Assisted Analysis of Biomedical Images
- SwarmFusion: Revolutionizing Disaster Response with Swarm Intelligence and Deep Learning
- OVANet: One-vs-All Network for Universal Domain Adaptation
- AI-GAs: AI-generating algorithms, an alternate paradigm for producing general artificial intelligence
- Segment This Thing: Foveated Tokenization for Efficient Point-Prompted Segmentation
- Superimage
- A Brief History of AI: How to Prevent Another Winter (A Critical Review)
- One-shot screening of potential peptide ligands on HR1 domain in COVID-19 glycosylated spike (S) protein with deep siamese network
- On the Ethics of Building AI in a Responsible Manner
- Exploring Structure Consistency for Deep Model Watermarking
- Are State-of-the-art Visual Place Recognition Techniques any Good for Aerial Robotics?
- Learning Orientation-Estimation Convolutional Neural Network for Building Detection in Optical Remote Sensing Image
- Adaptive Object Detection Using Adjacency and Zoom Prediction
- Deep Low-Shot Learning for Biological Image Classification and Visualization from Limited Training Samples
- DP-CGAN: Differentially Private Synthetic Data and Label Generation
- A survey of Object Classification and Detection based on 2D/3D data
- Transformer in Transformer
- Recurrent Neural Network from Adder's Perspective: Carry-lookahead RNN
- Effective and Efficient Dropout for Deep Convolutional Neural Networks
- Domain Generalization via Universal Non-volume Preserving Models
- Learning Spatial Fusion for Single-Shot Object Detection
- Towards Learning Convolutions from Scratch
- Convolutional Neural Networks for User Identificationbased on Motion Sensors Represented as Image
- Cross-X Learning for Fine-Grained Visual Categorization
- DropRegion Training of Inception Font Network for High-Performance Chinese Font Recognition
- A Global-Local Emebdding Module for Fashion Landmark Detection
- Bayesian Convolutional Neural Networks with Bernoulli Approximate Variational Inference
- Posterior Network: Uncertainty Estimation without OOD Samples via Density-Based Pseudo-Counts
- Probabilistic Discriminative Learning with Layered Graphical Models
- Semi-Supervised Hypothesis Transfer for Source-Free Domain Adaptation
- Deep Learning for Robust Motion Segmentation with Non-Static Cameras
- Deep Room Recognition Using Inaudible Echos
- An Unconstrained Layer-Peeled Perspective on Neural Collapse
- A Novel Training Protocol for Performance Predictors of Evolutionary Neural Architecture Search Algorithms
- MNEW: Multi-domain Neighborhood Embedding and Weighting for Sparse Point Clouds Segmentation
- Learning to Extend Program Graphs to Work-in-Progress Code
- CNN Profiler on Polar Coordinate Images for Tropical Cyclone Structure Analysis
- Dynamic Scale Inference by Entropy Minimization
- Towards a Unified Evaluation of Explanation Methods without Ground Truth
- Spectral Representations for Convolutional Neural Networks
- Gated CRF Loss for Weakly Supervised Semantic Image Segmentation
- SIFT Meets CNN: A Decade Survey of Instance Retrieval
- Imbalanced Image Classification with Complement Cross Entropy
- Towards Design Methodology of Efficient Fast Algorithms for Accelerating Generative Adversarial Networks on FPGAs
- Deep Learning for Rheumatoid Arthritis: Joint Detection and Damage\n Scoring in X-rays
- TGIF: A New Dataset and Benchmark on Animated GIF Description
- Self-Supervised Visual Representations for Cross-Modal Retrieval
- DR Loss: Improving Object Detection by Distributional Ranking
- Theory-based residual neural networks: A synergy of discrete choice models and deep neural networks
- Bidirectional Attention Network for Monocular Depth Estimation
- ATISS: Autoregressive Transformers for Indoor Scene Synthesis
- LaSO: Label-Set Operations networks for multi-label few-shot learning
- Adversarially Trained Deep Neural Semantic Hashing Scheme for Subjective Search in Fashion Inventory
- CNN-based Realized Covariance Matrix Forecasting
- Graph Neural Networks for Node-Level Predictions
- Deep Neural-Kernel Machines
- Segmentation and Generation of Magnetic Resonance Images by Deep Neural Networks
- Non-autoregressive Transformer by Position Learning
- Do Better ImageNet Models Transfer Better... for Image Recommendation?
- A Survey on Large-scale Machine Learning
- Universal Adversarial Perturbations: A Survey
- DL2: A Deep Learning-driven Scheduler for Deep Learning Clusters
- A concatenating framework of shortcut convolutional neural networks
- Reliable Identification of Redundant Kernels for Convolutional Neural Network Compression
- EQC : Ensembled Quantum Computing for Variational Quantum Algorithms
- Fashion Forward: Forecasting Visual Style in Fashion
- Learning Structures for Deep Neural Networks
- Leveraging Outdoor Webcams for Local Descriptor Learning
- Query Adaptive Late Fusion for Image Retrieval
- Inducing Hierarchical Compositional Model by Sparsifying Generator Network
- Image Retrieval on Real-life Images with Pre-trained Vision-and-Language Models
- Compact Deep Neural Networks for Computationally Efficient Gesture Classification From Electromyography Signals
- Data Augmentation for Deep Candlestick Learner
- Accumulated Polar Feature-based Deep Learning for Efficient and Lightweight Automatic Modulation Classification with Channel Compensation Mechanism
- Look at here : Utilizing supervision to attend subtle key regions
- All-You-Can-Fit 8-Bit Flexible Floating-Point Format for Accurate and Memory-Efficient Inference of Deep Neural Networks
- A Comprehensive Approach to Unsupervised Embedding Learning based on AND Algorithm
- Booster: An Accelerator for Gradient Boosting Decision Trees
- Towards Predicting the Likeability of Fashion Images
- Run, Don't Walk: Chasing Higher FLOPS for Faster Neural Networks
- Densely Connected Recurrent Residual (Dense R2UNet) Convolutional Neural Network for Segmentation of Lung CT Images
- Counting Fish and Dolphins in Sonar Images Using Deep Learning
- Deep neural networks can be improved using human-derived contextual expectations
- Who Will Share My Image? Predicting the Content Diffusion Path in Online Social Networks
- Discriminative Multi-modality Speech Recognition
- Big Data Goes Small: Real-Time Spectrum-Driven Embedded Wireless Networking Through Deep Learning in the RF Loop
- I-Nema: A Biological Image Dataset for Nematode Recognition
- Auto-encoding brain networks with applications to analyzing large-scale brain imaging datasets
- Differentiable Projection for Constrained Deep Learning
- DNNExplorer: A Framework for Modeling and Exploring a Novel Paradigm of FPGA-based DNN Accelerator
- RT3D: Achieving Real-Time Execution of 3D Convolutional Neural Networks on Mobile Devices
- PairNets: Novel Fast Shallow Artificial Neural Networks on Partitioned Subspaces
- Object-aware Feature Aggregation for Video Object Detection
- Self-supervised Visual Attribute Learning for Fashion Compatibility
- Universal Adder Neural Networks
- Machine Learning Applications for Therapeutic Tasks with Genomics Data
- A Preliminary Study on Data Augmentation of Deep Learning for Image Classification
- Comparing Attention-based Convolutional and Recurrent Neural Networks:\n Success and Limitations in Machine Reading Comprehension
- Distributed Averaging CNN-ELM for Big Data
- 2D versus 3D Convolutional Spiking Neural Networks Trained with Unsupervised STDP for Human Action Recognition
- Unsupervised Learning of Full-Waveform Inversion: Connecting CNN and Partial Differential Equation in a Loop
- Distributed stochastic optimization with large delays
- Pixel-Anchor: A Fast Oriented Scene Text Detector with Combined Networks
- Progressive Differentiable Architecture Search: Bridging the Depth Gap between Search and Evaluation
- A Large Scale Event-based Detection Dataset for Automotive
- Unsupervised Semantic Action Discovery from Video Collections
- Locally Enhanced Self-Attention: Combining Self-Attention and Convolution as Local and Context Terms
- Applying Tensor Decomposition to image for Robustness against Adversarial Attack
- Convolutional Neural Network on Semi-Regular Triangulated Meshes and its Application to Brain Image Data
- ConvPath: A Software Tool for Lung Adenocarcinoma Digital Pathological Image Analysis Aided by Convolutional Neural Network
- Artificial Intelligence for Pediatric Ophthalmology
- Learning Manifold Patch-Based Representations of Man-Made Shapes
- Linguistic Ordered Weighted Averaging based deep learning pooling for fault diagnosis in a wastewater treatment plant
- DeepLab: Semantic Image Segmentation with Deep Convolutional Nets, Atrous Convolution, and Fully Connected CRFs
- Deep Active Learning with Augmentation-based Consistency Estimation
- Group Ensemble: Learning an Ensemble of ConvNets in a single ConvNet
- GroupBERT: Enhanced Transformer Architecture with Efficient Grouped Structures
- Application of Facial Recognition using Convolutional Neural Networks for Entry Access Control
- Placepedia: Comprehensive Place Understanding with Multi-Faceted Annotations
- LyAm: Robust Non-Convex Optimization for Stable Learning in Noisy Environments
- StyleNAS: An Empirical Study of Neural Architecture Search to Uncover Surprisingly Fast End-to-End Universal Style Transfer Networks
- (Pen-) Ultimate DNN Pruning
- Learning a metacognition for object perception
- Uniform Generalization Bounds for Overparameterized Neural Networks
- Data-Dependent Randomized Smoothing
- Unsupervised Anomaly Detection for X-Ray Images
- Empirical Performance Analysis of Conventional Deep Learning Models for\n Recognition of Objects in 2-D Images
- A Transfer Learning Based Active Learning Framework for Brain Tumor Classification
- Improving Conversational Recommendation System by Pretraining on Billions Scale of Knowledge Graph
- A New Neural Network Architecture Invariant to the Action of Symmetry Subgroups
- Dependency Decomposition and a Reject Option for Explainable Models
- Towards localisation of keywords in speech using weak supervision
- Sewer-ML: A Multi-Label Sewer Defect Classification Dataset and Benchmark
- MEC: Memory-efficient Convolution for Deep Neural Network
- Cross-Lingual Vision-Language Navigation
- Try Harder: Hard Sample Generation and Learning for Clothes-Changing Person Re-ID
- A Comprehensive Survey for Real-World Industrial Defect Detection: Challenges, Approaches, and Prospects
- Functionalist Emotion Modeling in Biomimetic Reinforcement Learning
- SpaRTAN: Spatial Reinforcement Token-based Aggregation Network for Visual Recognition
- Predicting Deep Zero-Shot Convolutional Neural Networks using Textual Descriptions
- Event Specific Multimodal Pattern Mining with Image-Caption Pairs
- MARMOT: Masked Autoencoder for Modeling Transient Imaging
- Adaptive Object Detection with ESRGAN-Enhanced Resolution & Faster R-CNN
- Boundary Guidance Hierarchical Network for Real-Time Tongue Segmentation
- Measuring Robustness to Natural Distribution Shifts in Image Classification
- Network Space Search for Pareto-Efficient Spaces
- Modelling Urban Dynamics with Multi-Modal Graph Convolutional Networks
- A Novel Structured Natural Gradient Descent for Deep Learning
- Characterizing Inter-Layer Functional Mappings of Deep Learning Models
- Shelf-Supervised Mesh Prediction in the Wild
- OOWL500: Overcoming Dataset Collection Bias in the Wild
- CNN-based Density Estimation and Crowd Counting: A Survey
- Attention-Mechanism-based Tracking Method for Intelligent Internet of Vehicles
- Hashed Watermark as a Filter: Defeating Forging and Overwriting Attacks in Weight-based Neural Network Watermarking
- Multi-modal Face Pose Estimation with Multi-task Manifold Deep Learning
- Camera Style Adaptation for Person Re-identification
- HEIMDALL: a grapH-based sEIsMic Detector And Locator for microseismicity
- Semantic Image Cropping
- ABC-FHE : A Resource-Efficient Accelerator Enabling Bootstrappable Parameters for Client-Side Fully Homomorphic Encryption
- A Group Theoretic Analysis of the Symmetries Underlying Base Addition and Their Learnability by Neural Networks
- Deep Model Compression: Distilling Knowledge from Noisy Teachers
- Devanagari Handwritten Character Recognition using Convolutional Neural Network
- MFED: A System for Monitoring Family Eating Dynamics
- Matching Guided Distillation
- ContourRend: A Segmentation Method for Improving Contours by Rendering
- Deep Learning architectures for generalized immunofluorescence based nuclear image segmentation
- Deep Cross-Modal Hashing with Hashing Functions and Unified Hash Codes Jointly Learning
- Perception Improvement for Free: Exploring Imperceptible Black-box Adversarial Attacks on Image Classification
- Class-wise Dynamic Graph Convolution for Semantic Segmentation
- Lightweight image super-resolution with enhanced CNN
- Unsupervised Person Re-identification via Multi-Label Prediction and Classification based on Graph-Structural Insight
- GHFP: Gradually Hard Filter Pruning
- Ternary Residual Networks
- Transferring Styles for Reduced Texture Bias and Improved Robustness in Semantic Segmentation Networks
- Navigating the Challenges of AI-Generated Image Detection in the Wild: What Truly Matters?
- A Graph Sufficiency Perspective for Neural Networks
- Enhanced Few-shot Learning for Intrusion Detection in Railway Video Surveillance
- Convolutional Recurrent Residual U-Net Embedded with Attention Mechanism and Focal Tversky Loss Function for Cancerous Nuclei Detection
- LINTSRT: A Learning-driven Testbed for Intelligent Scheduling in Embedded Systems
- DARTS-ASR: Differentiable Architecture Search for Multilingual Speech Recognition and Adaptation
- Activity recognition from videos with parallel hypergraph matching on GPUs
- Augmentation Inside the Network
- Robust RGB-D Face Recognition Using Attribute-Aware Loss
- Superposed parameterised quantum circuits
- Towards the AlexNet Moment for Homomorphic Encryption: HCNN, theFirst Homomorphic CNN on Encrypted Data with GPUs
- Towards Reading Beyond Faces for Sparsity-Aware 4D Affect Recognition
- Label Noise Types and Their Effects on Deep Learning
- Collaborative Motion Prediction via Neural Motion Message Passing
- DeepLiDARFlow: A Deep Learning Architecture For Scene Flow Estimation Using Monocular Camera and Sparse LiDAR
- Bayesian neural networks and dimensionality reduction
- EENA: Efficient Evolution of Neural Architecture
- Polaritonic Machine Learning for Graph-based Data Analysis
- WebVision Challenge: Visual Learning and Understanding With Web Data
- Pre-training without Natural Images
- A Transfer Learning-Based Method for Water Body Segmentation in Remote Sensing Imagery: A Case Study of the Zhada Tulin Area
- BitParticle: Partializing Sparse Dual-Factors to Build Quasi-Synchronizing MAC Arrays for Energy-efficient DNNs
- Brain Stroke Detection and Classification Using CT Imaging with Transformer Models and Explainable AI
- QuarterMap: Efficient Post-Training Token Pruning for Visual State Space Models
- CAVA: A Visual Analytics System for Exploratory Columnar Data Augmentation Using Knowledge Graphs
- AWB-GCN: A Graph Convolutional Network Accelerator with Runtime Workload Rebalancing
- Classification of Industrial Control Systems screenshots using Transfer Learning
- Accelerating 2PC-based ML with Limited Trusted Hardware
- Modular network for high accuracy object detection
- Prior Knowledge about Attributes: Learning a More Effective Potential Space for Zero-Shot Recognition
- Recurrent Models of Visual Attention
- Using Deep Convolutional Neural Networks to Circumvent Morphological Feature Specification when Classifying Subvisible Protein Aggregates from Micro-Flow Images
- Active Subspace of Neural Networks: Structural Analysis and Universal Attacks
- DAA*: Deep Angular A Star for Image-based Path Planning
- MixNorm: Test-Time Adaptation Through Online Normalization Estimation
- A Binded VAE for Inorganic Material Generation
- Regret Bounds and Reinforcement Learning Exploration of EXP-based Algorithms
- AGCD-Net: Attention Guided Context Debiasing Network for Emotion Recognition
- Color Image Classification via Quaternion Principal Component Analysis Network
- A Meta Approach to Defend Noisy Labels by the Manifold Regularizer PSDR
- Modeling Partially Observed Nonlinear Dynamical Systems and Efficient Data Assimilation via Discrete-Time Conditional Gaussian Koopman Network
- Catastrophic Forgetting Mitigation Through Plateau Phase Activity Profiling
- Normalized vs Diplomatic Annotation: A Case Study of Automatic Information Extraction from Handwritten Uruguayan Birth Certificates
- SFedKD: Sequential Federated Learning with Discrepancy-Aware Multi-Teacher Knowledge Distillation
- Large-scale Bisample Learning on ID Versus Spot Face Recognition
- Towards Imperceptible JPEG Image Hiding: Multi-range Representations-driven Adversarial Stego Generation
- Weight-Sharing Neural Architecture Search: A Battle to Shrink the Optimization Gap
- Scene-Intuitive Agent for Remote Embodied Visual Grounding
- Oriented Objects as pairs of Middle Lines
- Perceptual Score: What Data Modalities Does Your Model Perceive?
- MUXConv: Information Multiplexing in Convolutional Neural Networks
- ResNEsts and DenseNEsts: Block-based DNN Models with Improved\n Representation Guarantees
- GPUHammer: Rowhammer Attacks on GPU Memories are Practical
- Physics-Informed Neural Networks with Hard Nonlinear Equality and Inequality Constraints
- PolyNet: A Pursuit of Structural Diversity in Very Deep Networks
- Multigranular Evaluation for Brain Visual Decoding
- Deriving Emotions and Sentiments from Visual Content: A Disaster Analysis Use Case
- Short-Term Temporal Convolutional Networks for Dynamic Hand Gesture Recognition
- Learning Pole Structures of Hadronic States using Predictive Uncertainty Estimation
- Temporal Unlearnable Examples: Preventing Personal Video Data from Unauthorized Exploitation by Object Tracking
- An Automated Classifier of Harmful Brain Activities for Clinical Usage Based on a Vision-Inspired Pre-trained Framework
- Kernel Quantization for Efficient Network Compression
- Understanding Dataset Bias in Medical Imaging: A Case Study on Chest X-rays
- Machine Learning based Post Processing Artifact Reduction in HEVC Intra Coding
- An Image Based Visual Servo Approach with Deep Learning for Robotic Manipulation
- Image Generation from Layout
- Bi-Classifier Determinacy Maximization for Unsupervised Domain Adaptation
- Distributed Inexact Successive Convex Approximation ADMM: Analysis-Part I
- Robust and Safe Traffic Sign Recognition using N-version with Weighted Voting
- Improved Hard Example Mining by Discovering Attribute-based Hard Person Identity
- Analyzing Adversarial Robustness of Deep Neural Networks in Pixel Space: a Semantic Perspective
- Aerial Maritime Vessel Detection and Identification
- Ancient Painting to Natural Image: A New Solution for Painting Processing
- M2-MFP: A Multi-Scale and Multi-Level Memory Failure Prediction Framework for Reliable Cloud Infrastructure
- Deep learning models for visibility forecasting using climatological data
- Deep Learning Techniques for Future Intelligent Cross-Media Retrieval
- SoftSignSGD(S3): An Enhanced Optimizer for Practical DNN Training and Loss Spikes Minimization Beyond Adam
- Energy-Efficient Supervised Learning with a Binary Stochastic Forward-Forward Algorithm
- HVI-CIDNet+: Beyond Extreme Darkness for Low-Light Image Enhancement
- Dual Mixup Regularized Learning for Adversarial Domain Adaptation
- Instant-Teaching: An End-to-End Semi-Supervised Object Detection Framework
- Centralized Copy-Paste: Enhanced Data Augmentation Strategy for Wildland Fire Semantic Segmentation
- Adversarial Defense Through Network Profiling Based Path Extraction
- Concept-Based Mechanistic Interpretability Using Structured Knowledge Graphs
- SoftReMish: A Novel Activation Function for Enhanced Convolutional Neural Networks for Visual Recognition Performance
- Contrastive and Transfer Learning for Effective Audio Fingerprinting through a Real-World Evaluation Protocol
- Generating private data with user customization
- FRAME Revisited: An Interpretation View Based on Particle Evolution
- Popcorn: Paillier Meets Compression For Efficient Oblivious Neural Network Inference
- ShadowNet: A Secure and Efficient On-device Model Inference System for Convolutional Neural Networks
- An Architecture Combining Convolutional Neural Network (CNN) and Support Vector Machine (SVM) for Image Classification
- Interpreting Shared Deep Learning Models via Explicable Boundary Trees
- Learning to Reconstruct and Segment 3D Objects
- Digital Twin: Values, Challenges and Enablers From a Modeling Perspective
- Training Very Deep Networks
- Self-Knowledge Distillation with Progressive Refinement of Targets
- AAG: Self-Supervised Representation Learning by Auxiliary Augmentation with GNT-Xent Loss
- Efficient Federated Learning with Timely Update Dissemination
- Deep Online Convex Optimization with Gated Games
- GASNet: Weakly-supervised Framework for COVID-19 Lesion Segmentation
- Investigating and Simplifying Masking-based Saliency Methods for Model\n Interpretability
- Batch Normalization with Enhanced Linear Transformation
- Recent Advance in Content-based Image Retrieval: A Literature Survey
- Simple Convergence Proof of Adam From a Sign-like Descent Perspective
- Hierarchical Predictive Coding Models in a Deep-Learning Framework
- Prediction of Tuberculosis using U-Net and segmentation techniques
- Generative Panoramic Image Stitching
- YOLO-APD: Enhancing YOLOv8 for Robust Pedestrian Detection on Complex Road Geometries
- LAID: Lightweight AI-Generated Image Detection in Spatial and Spectral Domains
- Efficient Computation Reduction in Bayesian Neural Networks Through Feature Decomposition and Memorization
- Domain Adaptive Object Detection via Feature Separation and Alignment
- LVM4CSI: Enabling Direct Application of Pre-Trained Large Vision Models for Wireless Channel Tasks
- Effort-Optimized, Accuracy-Driven Labelling and Validation of Test Inputs for DL Systems: A Mixed-Integer Linear Programming Approach
- Towards Single Stage Weakly Supervised Semantic Segmentation
- Improving Adversarial Robustness in Weight-quantized Neural Networks
- Visual Identification of Individual Holstein-Friesian Cattle via Deep Metric Learning
- Channel Distillation: Channel-Wise Attention for Knowledge Distillation
- Discriminative Localized Sparse Representations for Breast Cancer Screening
- Salient Object Detection Combining a Self-attention Module and a Feature Pyramid Network
- Neural Velocity for hyperparameter tuning
- Model Compression using Progressive Channel Pruning
- Constructive Universal Approximation and Sure Convergence for Multi-Layer Neural Networks
- CP-Dilatation: A Copy-and-Paste Augmentation Method for Preserving the Boundary Context Information of Histopathology Images
- Deep learning for caries detection: A systematic review
- Large-scale Continuous Gesture Recognition Using Convolutional Neural Networks
- Cyclic Differentiable Architecture Search
- PRING: Rethinking Protein-Protein Interaction Prediction from Pairs to Graphs
- Learning by Examples Based on Multi-level Optimization
- An Experimental Study of Deep Convolutional Features For Iris Recognition
- The Replica Dataset: A Digital Replica of Indoor Spaces
- Adversarial Examples Improve Image Recognition
- U-Net: Convolutional Networks for Biomedical Image Segmentation
- Novel Adaptive Binary Search Strategy-First Hybrid Pyramid- and Clustering-Based CNN Filter Pruning Method without Parameters Setting
- Distribution Mismatch Correction for Improved Robustness in Deep Neural Networks
- Image recognition via Vietoris-Rips complex
- Optimizing Neural Network for Computer Vision task in Edge Device
- Multiview convolutional neural networks for lung nodule classification
- Stochastic Mirror Descent on Overparameterized Nonlinear Models: Convergence, Implicit Regularization, and Generalization
- Delving into VoxCeleb: environment invariant speaker recognition
- Understanding Modern Techniques in Optimization: Frank-Wolfe, Nesterov's Momentum, and Polyak's Momentum
- Graph Convolutional Networks with EigenPooling
- ConvNets vs. Transformers: Whose Visual Representations are More Transferable?
- Structured Convolution Matrices for Energy-efficient Deep learning
- Thousand-Brains Systems: Sensorimotor Intelligence for Rapid, Robust Learning and Inference
- Zero-shot Adversarial Quantization
- Improving out-of-distribution generalization via multi-task self-supervised pretraining
- Variational Bayesian Dropout with a Hierarchical Prior
- Growing Efficient Deep Networks by Structured Continuous Sparsification
- Centralized Information Interaction for Salient Object Detection
- Stochastic Optimization with Laggard Data Pipelines
- Temporal Contrastive Graph Learning for Video Action Recognition and Retrieval
- Early Convolutions Help Transformers See Better
- The Scattering Compositional Learner: Discovering Objects, Attributes, Relationships in Analogical Reasoning
- SymNet: Symmetrical Filters in Convolutional Neural Networks
- Overcoming Statistical Shortcuts for Open-ended Visual Counting
- ProductNet: a Collection of High-Quality Datasets for Product Representation Learning
- PromptSR: Cascade Prompting for Lightweight Image Super-Resolution
- Teaching AI to Explain its Decisions Using Embeddings and Multi-Task Learning
- SkeletonNet: A Topology-Preserving Solution for Learning Mesh Reconstruction of Object Surfaces from RGB Images
- Overcoming the low signal-to-noise problem for hybrid mode-selective photonic lantern-based wavefront correction using machine learning
- SemanticAdv: Generating Adversarial Examples via Attribute-conditional Image Editing
- On Class Imbalance and Background Filtering in Visual Relationship Detection
- Search for spatial coincidences between galaxy mergers and Fermi-LAT 4FGL-DR4 sources
- Node-By-Node Greedy Deep Learning for Interpretable Features
- Dynamic Reliability Management in Neuromorphic Computing
- Gaussian-LIC2: LiDAR-Inertial-Camera Gaussian Splatting SLAM
- Chameleon: A Semi-AutoML framework targeting quick and scalable development and deployment of production-ready ML systems for SMEs
- Incremental Training and Group Convolution Pruning for Runtime DNN Performance Scaling on Heterogeneous Embedded Platforms
- Adaptive Test-Time Augmentation for Low-Power CPU
- CoDR: Computation and Data Reuse Aware CNN Accelerator
- Towards Analysis-friendly Face Representation with Scalable Feature and Texture Compression
- Active Screening for Recurrent Diseases: A Reinforcement Learning Approach
- 3D PixBrush: Image-Guided Local Texture Synthesis
- Mutual Graph Learning for Camouflaged Object Detection
- Shakeout: A New Approach to Regularized Deep Neural Network Training
- Multi-Level Fusion Graph Neural Network for Molecule Property Prediction
- Adversarial Attacks on Deep Learning Models in Natural Language Processing: A Survey
- C-MIL: Continuation Multiple Instance Learning for Weakly Supervised Object Detection
- Backward-Compatible Prediction Updates: A Probabilistic Approach
- Global Variational Inference Enhanced Robust Domain Adaptation
- A CNN–LSTM model for gold price time-series forecasting
- Medical Datasets Collections for Artificial Intelligence-based Medical Image Analysis
- Convolutional Kernel Networks
- Compression and Interpretability of Deep Neural Networks via Tucker Tensor Layer: From First Principles to Tensor Valued Back-Propagation
- Unsupervised Deep Representation Learning for Real-Time Tracking
- TransGAN: Two Pure Transformers Can Make One Strong GAN, and That Can Scale Up
- A Brief Survey and an Application of Semantic Image Segmentation for Autonomous Driving
- Do ImageNet Classifiers Generalize to ImageNet?
- PAD-Net: Multi-Tasks Guided Prediction-and-Distillation Network for Simultaneous Depth Estimation and Scene Parsing
- HarDNet: A Low Memory Traffic Network
- Expanding phenological insights: automated phenostage annotation with community science plant images
- Inverse Synthetic Aperture Fourier Ptychography
- Ensemble of Deep Learned Features for Melanoma Classification
- PoTrojan: powerful neural-level trojan designs in deep learning models
- Depth Evaluation for Metal Surface Defects by Eddy Current Testing using Deep Residual Convolutional Neural Networks
- SCOPE: Scalable Composite Optimization for Learning on Spark
- Gabor filter incorporated CNN for compression
- Characterizing Compute-Communication Overlap in GPU-Accelerated Distributed Deep Learning: Performance and Power Implications
- Linear Attention with Global Context: A Multipole Attention Mechanism for Vision and Physics
- Automated Grading of Students' Handwritten Graphs: A Comparison of Meta-Learning and Vision-Large Language Models
- Position: A Theory of Deep Learning Must Include Compositional Sparsity
- Stacked Convolutional Neural Network for Diagnosis of COVID-19 Disease from X-ray Images
- FNA++: Fast Network Adaptation via Parameter Remapping and Architecture Search
- Hierarchically Compositional Tasks and Deep Convolutional Networks
- Optimisation Is Not What You Need
- Statistical Inference for Stochastic Gradient Descent: Beyond Finite Variance
- LMPNet for Weakly-supervised Keypoint Discovery
- Tomato plant disease detection using transfer learning with C-GAN synthetic images
- C2C-GenDA: Cluster-to-Cluster Generation for Data Augmentation of Slot Filling
- Robust Real-Time Pedestrian Detection on Embedded Devices
- Cross-Domain Image Classification through Neural-Style Transfer Data Augmentation
- The Book of Life approach: Enabling richness and scale for life course research
- Can Artificial Intelligence solve the blockchain oracle problem? Unpacking the Challenges and Possibilities
- Selective Feature Re-Encoded Quantum Convolutional Neural Network with Joint Optimization for Image Classification
- Towards Compact CNNs via Collaborative Compression
- Decamouflage: A Framework to Detect Image-Scaling Attacks on Convolutional Neural Networks
- Semi-Supervising Learning, Transfer Learning, and Knowledge Distillation with SimCLR
- Communication-Computation Trade-Off in Resource-Constrained Edge Inference
- Quantum reinforcement learning in dynamic environments
- Robust Local Features for Improving the Generalization of Adversarial Training
- Spectral Metric for Dataset Complexity Assessment
- Contextual Multi-Scale Region Convolutional 3D Network for Activity Detection
- Contact Pose Identification for Peg-in-Hole Assembly under Uncertainties
- JSR-Net: A Deep Network for Joint Spatial-Radon Domain CT Reconstruction from incomplete data
- Gradient Short-Circuit: Efficient Out-of-Distribution Detection via Feature Intervention
- POST: Photonic Swin Transformer for Automated and Efficient Prediction of PCSEL
- Beyond Categorical Label Representations for Image Classification
- Adversarial Item Promotion: Vulnerabilities at the Core of Top-N Recommenders that Use Images to Address Cold Start
- Trigger Hunting with a Topological Prior for Trojan Detection
- TFLMS: Large Model Support in TensorFlow by Graph Rewriting
- Complex network prediction using deep learning
- evMLP: An Efficient Event-Driven MLP Architecture for Vision
- Progressive DARTS: Bridging the Optimization Gap for NAS in the Wild
- BlockSwap: Fisher-guided Block Substitution for Network Compression on a Budget
- Long-Tailed Distribution-Aware Router For Mixture-of-Experts in Large Vision-Language Model
- Deep Learning in Spiking Neural Networks
- Tangma: A Tanh-Guided Activation Function with Learnable Parameters
- Towards a Domain Specific Solution for a New Generation of Wireless Modems
- A Review on Sound Source Localization in Robotics: Focusing on Deep Learning Methods
- Complexity of Representation and Inference in Compositional Models with Part Sharing
- Preconditioned Stochastic Gradient Langevin Dynamics for Deep Neural Networks
- BA-Net: Dense Bundle Adjustment Network
- AI-Generated Video Detection via Perceptual Straightening
- Visual Anagrams Reveal Hidden Differences in Holistic Shape Processing Across Vision Models
- A deep learning-based framework for an automated defect detection system for sewer pipes
- Underground sewer pipe condition assessment based on convolutional neural networks
- Learning Dense Feature Matching via Lifting Single 2D Image to 3D Space
- Combining Diverse Feature Priors
- Regularizing Neural Networks via Minimizing Hyperspherical Energy
- Unsupervised Representation Learning for Binary Networks by Joint Classifier Learning
- X-LineNet: Detecting Aircraft in Remote Sensing Images by a pair of Intersecting Line Segments
- SPARK: Spatial-aware Online Incremental Attack Against Visual Tracking
- Multi-modal Visual Tracking: Review and Experimental Comparison
- Adversarial Examples, Uncertainty, and Transfer Testing Robustness in Gaussian Process Hybrid Deep Networks
- Discriminative Local Sparse Representation by Robust Adaptive Dictionary Pair Learning
- DeepLab: Semantic Image Segmentation with Deep Convolutional Nets, Atrous Convolution, and Fully Connected CRFs
- Real-time Faulted Line Localization and PMU Placement in Power Systems through Convolutional Neural Networks
- Towards End-to-End Deep Learning for Autonomous Racing: On Data Collection and a Unified Architecture for Steering and Throttle Prediction
- Semi-Supervised Domain Adaptation via Selective Pseudo Labeling and Progressive Self-Training
- Consensus-based optimization for closed-box adversarial attacks and a connection to evolution strategies
- Structured Mechanical Models for Robot Learning and Control
- Automatically Learning Data Augmentation Policies for Dialogue Tasks
- Segmented Operations using Matrix Multiplications
- QPART: Adaptive Model Quantization and Dynamic Workload Balancing for Accuracy-aware Edge Inference
- A Survey on Deep Learning for Multimodal Data Fusion
- Pruning by Block Benefit: Exploring the Properties of Vision Transformer Blocks during Domain Adaptation
- SkeleMotion: A New Representation of Skeleton Joint Sequences Based on Motion Information for 3D Action Recognition
- Faster and Accurate Classification for JPEG2000 Compressed Images in Networked Applications
- Both Asymptotic and Non-Asymptotic Convergence of Quasi-Hyperbolic Momentum using Increasing Batch Size
- GViT: Representing Images as Gaussians for Visual Recognition
- GradEscape: A Gradient-Based Evader Against AI-Generated Text Detectors
- Optimization of Graph Neural Networks with Natural Gradient Descent
- Data Augmentation Revisited: Rethinking the Distribution Gap between Clean and Augmented Data
- Multi-Scale RCNN Model for Financial Time-series Classification
- Is Robustness the Cost of Accuracy? -- A Comprehensive Study on the Robustness of 18 Deep Image Classification Models
- Spurious-Aware Prototype Refinement for Reliable Out-of-Distribution Detection
- Safe Reinforcement Learning via Probabilistic Shields
- CuRe: Cultural Gaps in the Long Tail of Text-to-Image Systems
- Semantic Attention and Scale Complementary Network for Instance Segmentation in Remote Sensing Images
- OGNet: Salient Object Detection with Output-guided Attention Module
- Competitive Distillation: A Simple Learning Strategy for Improving Visual Classification
- BAPE: Learning an Explicit Bayes Classifier for Long-tailed Visual Recognition
- Multi-focus Attention Network for Efficient Deep Reinforcement Learning
- Class-Balanced Loss Based on Effective Number of Samples
- Can WiFi Estimate Person Pose?
- Characterizing the Deep Neural Networks Inference Performance of Mobile Applications
- Quantitative Propagation of Chaos for SGD in Wide Neural Networks
- CAMERAS: Enhanced Resolution And Sanity preserving Class Activation Mapping for image saliency
- Exploiting Linear Structure Within Convolutional Networks for Efficient Evaluation
- Digital phase-only holography using deep conditional generative models
- Learning Obstacle Representations for Neural Motion Planning
- CNN based texture synthesize with Semantic segment
- A Survey of Human-in-the-loop for Machine Learning
- On the adequacy of untuned warmup for adaptive optimization
- Multi-Object Classification and Unsupervised Scene Understanding Using Deep Learning Features and Latent Tree Probabilistic Models
- Using learned optimizers to make models robust to input noise
- Joint Semantic Domain Alignment and Target Classifier Learning for Unsupervised Domain Adaptation
- Exact Adversarial Attack to Image Captioning via Structured Output Learning with Latent Variables
- Transductive Episodic-Wise Adaptive Metric for Few-Shot Learning
- PipeNet: Selective Modal Pipeline of Fusion Network for Multi-Modal Face Anti-Spoofing
- Semantic Amodal Segmentation
- Unsupervised Audiovisual Synthesis via Exemplar Autoencoders
- A Microprocessor implemented in 65nm CMOS with Configurable and Bit-scalable Accelerator for Programmable In-memory Computing
- Persistence Paradox in Dynamic Science
- Tool as Embodiment for Recursive Manipulation
- Crater Detection via Convolutional Neural Networks
- Weakly-supervised Disentangling with Recurrent Transformations for 3D View Synthesis
- D2C: Diffusion-Denoising Models for Few-shot Conditional Generation
- Revolutionizing gastroenterology and hepatology with artificial intelligence: From precision diagnosis to equitable healthcare through interdisciplinary practice
- LDMI: An Information-theoretic Noise-robust Loss Function
- Topmetal CMOS direct charge sensing plane for neutrinoless double-beta decay search in high-pressure gaseous TPC
- A SOT-MRAM-based Processing-In-Memory Engine for Highly Compressed DNN Implementation
- Real-time Speech Frequency Bandwidth Extension
- Improving Token-based Object Detection with Video
- Optimization of Wireless Sensor Network Deployment for Spatiotemporal Reconstruction and Prediction
- Towards Distributed Neural Architectures
- ADMM-Net: A Deep Learning Approach for Compressive Sensing MRI
- One Video to Steal Them All: 3D-Printing IP Theft through Optical Side-Channels
- SpikeSMOKE: Spiking Neural Networks for Monocular 3D Object Detection with Cross-Scale Gated Coding
- Pixels-to-Graph: Real-time Integration of Building Information Models and Scene Graphs for Semantic-Geometric Human-Robot Understanding
- Advancing Facial Stylization through Semantic Preservation Constraint and Pseudo-Paired Supervision
- Detecting Interspecific Positive Selection Using Convolutional Neural Networks
- Expressivity of Neural Networks via Chaotic Itineraries beyond Sharkovsky's Theorem
- ProARD: progressive adversarial robustness distillation: provide wide range of robust students
- UniCA: Unified Covariate Adaptation for Time Series Foundation Model
- Deep RNN Framework for Visual Sequential Applications
- Recursive Binary Neural Network Learning Model with 2.28b/Weight Storage Requirement
- Learning to adapt class-specific features across domains for semantic segmentation
- Finding Task-Relevant Features for Few-Shot Learning by Category Traversal
- Application of quantum machine learning using variational quantum classifier in accelerator physics
- North American extreme temperature events and related large scale meteorological patterns: a review of statistical methods, dynamics, modeling, and trends
- An Unsupervised Deep-Learning Method for Fingerprint Classification: the CCAE Network and the Hybrid Clustering Strategy
- Robust Out-of-Distribution Detection on Deep Probabilistic Generative Models
- Feature Pyramid and Hierarchical Boosting Network for Pavement Crack Detection
- KuraNet: Systems of Coupled Oscillators that Learn to Synchronize
- Learn to Predict Vertical Track Irregularity with Extremely Imbalanced Data
- rQdia: Regularizing Q-Value Distributions With Image Augmentation
- Towards Leveraging End-of-Life Tools as an Asset: Value Co-Creation based on Deep Learning in the Machining Industry
- Stochastic and Non-local Closure Modeling for Nonlinear Dynamical Systems via Latent Score-based Generative Models
- Very simple statistical evidence that AlphaGo has exceeded human limits\n in playing GO game
- Lightweight Multi-Frame Integration for Robust YOLO Object Detection in Videos
- Efficient Certified Reasoning for Binarized Neural Networks
- ACNN: a Full Resolution DCNN for Medical Image Segmentation
- Effectiveness of Optimization Algorithms in Deep Image Classification
- The duality structure gradient descent algorithm: analysis and\n applications to neural networks
- Curvature Enhanced Data Augmentation for Regression
- Control and optimization for Neural Partial Differential Equations in Supervised Learning
- SFNet: Fusion of Spatial and Frequency-Domain Features for Remote Sensing Image Forgery Detection
- Machine-Learning-Assisted Photonic Device Development: A Multiscale Approach from Theory to Characterization
- Memory Enhanced Global-Local Aggregation for Video Object Detection
- Ark: An Open-source Python-based Framework for Robot Learning
- Distillation Guided Residual Learning for Binary Convolutional Neural Networks
- Curating art exhibitions using machine learning
- A Survey of LLM-Driven AI Agent Communication: Protocols, Security Risks, and Defense Countermeasures
- Known-class Aware Self-ensemble for Open Set Domain Adaptation
- using multiple losses for accurate facial age estimation
- Generative model for optimal density estimation on unknown manifold
- PrivacyXray: Detecting Privacy Breaches in LLMs through Semantic Consistency and Probability Certainty
- Deceiving End-to-End Deep Learning Malware Detectors using Adversarial Examples
- Comparative Performance of Finetuned ImageNet Pre-trained Models for Electronic Component Classification
- PAI-BPR: Personalized Outfit Recommendation Scheme with Attribute-wise Interpretability
- A critical analysis of self-supervision, or what we can learn from a single image
- Cephalometric Landmark Detection by AttentiveFeature Pyramid Fusion and Regression-Voting
- Intrinsic-Extrinsic Convolution and Pooling for Learning on 3D Protein\n Structures
- On the Demystification of Knowledge Distillation: A Residual Network Perspective
- Deep Multi-scale Discriminative Networks for Double JPEG Compression Forensics
- Identification of hydrodynamic instability by convolutional neural networks
- STA: Adversarial Attacks on Siamese Trackers
- A Comprehensive Survey on Transfer Learning
- Anomaly Detection with Prototype-Guided Discriminative Latent Embeddings
- Neuro-Symbolic Execution: The Feasibility of an Inductive Approach to Symbolic Execution
- Sequence-to-Sequence Data Augmentation for Dialogue Language Understanding
- Partially Observable Residual Reinforcement Learning for PV-Inverter-Based Voltage Control in Distribution Grids
- Emotion Detection on User Front-Facing App Interfaces for Enhanced Schedule Optimization: A Machine Learning Approach
- SNN: Stacked Neural Networks
- Augmenting C. elegans Microscopic Dataset for Accelerated Pattern Recognition
- SiamEvent: Event-based Object Tracking via Edge-aware Similarity Learning with Siamese Networks
- Transferable Semantic Augmentation for Domain Adaptation
- Novelty-Prepared Few-Shot Classification
- SIM-Net: A Multimodal Fusion Network Using Inferred 3D Object Shape Point Clouds from RGB Images for 2D Classification
- BulletGen: Improving 4D Reconstruction with Bullet-Time Generation
- Biologically-Motivated Deep Learning Method using Hierarchical Competitive Learning
- Generalizing vision-language models to novel domains: A comprehensive survey
- Weakly-supervised Generative Adversarial Networks for medical image classification
- A Deep Convolutional Neural Network-Based Novel Class Balancing for Imbalance Data Segmentation
- DIP: Unsupervised Dense In-Context Post-training of Visual Representations
- Data Dwarfs: A Lens Towards Fully Understanding Big Data and AI Workloads
- Fast AI Model Splitting over Edge Networks
- Extreme Value Preserving Networks
- Dual-Forward Path Teacher Knowledge Distillation: Bridging the Capacity Gap Between Teacher and Student
- Exploiting Lightweight Hierarchical ViT and Dynamic Framework for Efficient Visual Tracking
- What Matters in Learning from Offline Human Demonstrations for Robot Manipulation
- Expressive TTS Training with Frame and Style Reconstruction Loss
- Witches' Brew: Industrial Scale Data Poisoning via Gradient Matching
- Scale-Invariant Convolutional Neural Networks
- Advancing Talking Head Generation: A Comprehensive Survey of Multi-Modal Methodologies, Datasets, Evaluation Metrics, and Loss Functions
- Facial Age Estimation using Convolutional Neural Networks
- A hierarchical deep learning framework for the consistent classification of land use objects in geospatial databases
- Advanced Applications of Generative AI in Actuarial Science: Case Studies Beyond ChatGPT
- Fine-Tuned Vision Transformers Capture Complex Wheat Spike Morphology for Volume Estimation from RGB Images
- Splitformer: An improved early-exit architecture for automatic speech recognition on edge devices
- Fully Learnable Group Convolution for Acceleration of Deep Neural Networks
- Assume, Augment and Learn: Unsupervised Few-Shot Meta-Learning via Random Labels and Data Augmentation
- 4DGT: Learning a 4D Gaussian Transformer Using Real-World Monocular Videos
- On the Robustness of Human-Object Interaction Detection against Distribution Shift
- Fast Incremental Learning for Off-Road Robot Navigation
- Generative NeuroEvolution for Deep Learning
- My First Deep Learning System of 1991 + Deep Learning Timeline 1962-2013
- Classification of Tents in Street Bazaars Using CNN
- DRO-Augment Framework: Robustness by Synergizing Wasserstein Distributionally Robust Optimization and Data Augmentation
- TEST: an End-to-End Network Traffic Examination and Identification Framework Based on Spatio-Temporal Features Extraction
- Detecting aquatic plant transmission on floatplanes using computer vision
- Visual Tracking via Shallow and Deep Collaborative Model
- Sketch2code: Generating a website from a paper mockup
- DSAM: A Distance Shrinking with Angular Marginalizing Loss for High Performance Vehicle Re-identificatio
- Right whale recognition using convolutional neural networks
- VeriLocc: End-to-End Cross-Architecture Register Allocation via LLM
- AQUA20: A Benchmark Dataset for Underwater Species Classification under Challenging Conditions
- Understanding and Testing Generalization of Deep Networks on Out-of-Distribution Data
- Differentiable Multiple Shooting Layers
- Sequence-to-Sequence Models with Attention Mechanistically Map to the Architecture of Human Memory Search
- Explore the vulnerability of black-box models via diffusion models
- RoIMix: Proposal-Fusion among Multiple Images for Underwater Object Detection
- Show, Match and Segment: Joint Weakly Supervised Learning of Semantic Matching and Object Co-segmentation
- On Feature Normalization and Data Augmentation
- Thermometry of simulated Bose--Einstein condensates using machine learning
- A Deep Learning Based Fast Image Saliency Detection Algorithm
- Lookup Table-based Multiplication-free All-digital DNN Accelerator Featuring Self-Synchronous Pipeline Accumulation
- Off-Policy Actor-Critic for Adversarial Observation Robustness: Virtual Alternative Training via Symmetric Policy Evaluation
- Population Based Augmentation: Efficient Learning of Augmentation Policy Schedules
- Totally Deep Support Vector Machines
- Aligning where to see and what to tell: image caption with region-based attention and scene factorization
- LeanResNet: A Low-cost Yet Effective Convolutional Residual Networks
- TAPESTRY: A Blockchain based Service for Trusted Interaction Online
- Adversarial Dual Distinct Classifiers for Unsupervised Domain Adaptation
- VRFP: On-the-fly Video Retrieval using Web Images and Fast Fisher Vector Products
- RobustART: Benchmarking Robustness on Architecture Design and Training Techniques
- Fast and Accurate, Convolutional Neural Network Based Approach for Object Detection from UAV
- Pixel DAG-Recurrent Neural Network for Spectral-Spatial Hyperspectral Image Classification
- Multiscale Adaptive Representation of Signals: I. The Basic Framework
- Deep Learning in Ultrasound Elastography Imaging
- A Brain-to-Population Graph Learning Framework for Diagnosing Brain Disorders
- Noise Fusion-based Distillation Learning for Anomaly Detection in Complex Industrial Environments
- TreeNet: A lightweight One-Shot Aggregation Convolutional Network
- Development of the algorithm for differentiating bone metastases and trauma of the ribs in bone scintigraphy and demonstration of visual evidence of the algorithm -- Using only anterior bone scan view of thorax
- Scheduled Differentiable Architecture Search for Visual Recognition
- Itsy Bitsy SpiderNet: Fully Connected Residual Network for Fraud Detection
- Turning old models fashion again: Recycling classical CNN networks using the Lattice Transformation
- Pay Attention to MLPs
- Reactive Transport Modeling with Physics-Informed Machine Learning for Critical Minerals Applications
- ControlVLA: Few-shot Object-centric Adaptation for Pre-trained Vision-Language-Action Models
- Reinforcement Learning with Augmented Data
- Translate-to-Recognize Networks for RGB-D Scene Recognition
- Layer Dynamics of Linearised Neural Nets
- Prototype Rectification for Few-Shot Learning
- Deep Stable Learning for Out-Of-Distribution Generalization
- Classification of Multi-Parametric Body MRI Series Using Deep Learning
- NTIRE 2025 Image Shadow Removal Challenge Report
- Real-Time Video Highlights for Yahoo Esports
- Image Captioning Based on a Hierarchical Attention Mechanism and Policy Gradient Optimization
- Rethinking Semantic Segmentation Evaluation for Explainability and Model Selection
- Alignment between Brains and AI: Evidence for Convergent Evolution across Modalities, Scales and Training Trajectories
- Out-of-Band Modality Synergy Based Multi-User Beam Prediction and Proactive BS Selection with Zero Pilot Overhead
- Diffusion-based Counterfactual Augmentation: Towards Robust and Interpretable Knee Osteoarthritis Grading
- Self-Supervised Learning via Conditional Motion Propagation
- Analytical Moment Regularizer for Gaussian Robust Networks
- LLC: Accurate, Multi-purpose Learnt Low-dimensional Binary Codes
- A Tunable Robust Pruning Framework Through Dynamic Network Rewiring of DNNs
- Learning Graph Neural Networks with Approximate Gradient Descent
- NERO: Explainable Out-of-Distribution Detection with Neuron-level Relevance
- Partitioning for Intrinsic Model Inversion Resistance in Collaborative Inference
- HiPreNets: High-Precision Neural Networks through Progressive Training
- Demonstrating Superresolution in Radar Range Estimation Using a Denoising Autoencoder
- Attentional Neural Network: Feature Selection Using Cognitive Feedback
- JSRT: James-Stein Regression Tree
- Heart rate and respiratory rate prediction from noisy real-world smartphone based on Deep Learning methods
- Exploring Temporal Information for Improved Video Understanding
- Vehicle Detection in Deep Learning
- Supervised Online Hashing via Similarity Distribution Learning
- Towards Automatic Construction of Diverse, High-quality Image Dataset
- Foundation Model Insights and a Multi-Model Approach for Superior Fine-Grained One-shot Subset Selection
- A Hybrid Method for Traffic Flow Forecasting Using Multimodal Deep Learning
- Peering into the Unknown: Active View Selection with Neural Uncertainty Maps for 3D Reconstruction
- Direct tensor processing with coherent light
- Self-supervised Representation Learning with Local Aggregation for Image-based Profiling
- Universal rates of ERM for agnostic learning
- Image Segmentation with Large Language Models: A Survey with Perspectives for Intelligent Transportation Systems
- ResNets Are Deeper Than You Think
- Rethinking Text Segmentation: A Novel Dataset and A Text-Specific Refinement Approach
- A Deep Learning Bidirectional Temporal Tracking Algorithm for Automated Blood Cell Counting from Non-invasive Capillaroscopy Videos
- Learning Novel Objects Continually Through Curiosity
- Academic Performance Estimation with Attention-based Graph Convolutional Networks
- Fisher Discriminative Least Squares Regression for Image Classification
- Local Area Transform for Cross-Modality Correspondence Matching and Deep Scene Recognition
- Rigor in AI: Doing Rigorous AI Work Requires a Broader, Responsible AI-Informed Conception of Rigor
- Busting the Paper Ballot: Voting Meets Adversarial Machine Learning
- Bilinear Supervised Hashing Based on 2D Image Features
- From Tool Calling to Symbolic Thinking: LLMs in a Persistent Lisp Metaprogramming Loop
- GMT: General Motion Tracking for Humanoid Whole-Body Control
- A Hierarchical Test Platform for Vision Language Model (VLM)-Integrated Real-World Autonomous Driving
- Meta-SurDiff: Classification Diffusion Model Optimized by Meta Learning is Reliable for Online Surgical Phase Recognition
- Embedding physical symmetries into machine-learned reduced plasma physics models via data augmentation
- Theoretically Unmasking Inference Attacks Against LDP-Protected Clients in Federated Vision Models
- The Care Label Concept: A Certification Suite for Trustworthy and Resource-Aware Machine Learning
- Consumer Image Quality Prediction using Recurrent Neural Networks for Spatial Pooling
- Rapid Neural Architecture Search by Learning to Generate Graphs from Datasets
- Diagnosing and Improving Diffusion Models by Estimating the Optimal Loss Value
- Deep Kernel Learning via Random Fourier Features
- Convolutional Neural Networks with Dynamic Regularization
- Advancing Image-Based Grapevine Variety Classification with a New Benchmark and Evaluation of Masked Autoencoders
- PhenoKG: Knowledge Graph-Driven Gene Discovery and Patient Insights from Phenotypes Alone
- Curriculum By Smoothing
- A Survey on World Models Grounded in Acoustic Physical Information
- Evolution of ReID: From Early Methods to LLM Integration
- A Unified Joint Maximum Mean Discrepancy for Domain Adaptation
- A Multi-Scale Mapping Approach Based on a Deep Learning CNN Model for Reconstructing High-Resolution Urban DEMs
- Active Crowd Counting with Limited Supervision
- MeliusNet: Can Binary Neural Networks Achieve MobileNet-level Accuracy?
- An Augmented Transformer Architecture for Natural Language Generation Tasks
- Do They All Look the Same? Deciphering Chinese, Japanese and Koreans by Fine-Grained Deep Learning
- CSI2Image: Image Reconstruction from Channel State Information Using Generative Adversarial Networks
- Linear Range in Gradient Descent
- DiffGCN: Graph Convolutional Networks via Differential Operators and Algebraic Multigrid Pooling
- Statistics of Deep Generated Images
- Depth Not Needed - An Evaluation of RGB-D Feature Encodings for Off-Road Scene Understanding by Convolutional Neural Network
- Bring Your Own Codegen to Deep Learning Compiler
- FDNet: A Deep Learning Approach with Two Parallel Cross Encoding Pathways for Precipitation Nowcasting
- Faster and Simpler Siamese Network for Single Object Tracking
- ConTNet: Why not use convolution and transformer at the same time?
- A Survey on Deep Domain Adaptation and Tiny Object Detection Challenges, Techniques and Datasets
- Stagewise Training Accelerates Convergence of Testing Error Over SGD
- Geometric Deep Learning: Grids, Groups, Graphs, Geodesics, and Gauges
- Binarized Weight Error Networks With a Transition Regularization Term
- Precision Gating: Improving Neural Network Efficiency with Dynamic Dual-Precision Activations
- Deep Learning-Based Automated Image Segmentation for Concrete Petrographic Analysis
- Self-Supervision and Spatial-Sequential Attention Based Loss for Multi-Person Pose Estimation
- Towards Interpretable Adversarial Examples via Sparse Adversarial Attack
- Instance Credibility Inference for Few-Shot Learning
- CoSA: Scheduling by Constrained Optimization for Spatial Accelerators
- Poisoning the Search Space in Neural Architecture Search
- Provable Bounds for Learning Some Deep Representations
- Dynamic Graph CNN for Learning on Point Clouds
- Fingerspelling recognition in the wild with iterative visual attention
- Towards Unified INT8 Training for Convolutional Neural Network
- Deep-LK for Efficient Adaptive Object Tracking
- DMV: Visual Object Tracking via Part-level Dense Memory and Voting-based Retrieval
- Graph Convolutional Network for Recommendation with Low-pass Collaborative Filters
- Connecting Touch and Vision via Cross-Modal Prediction
- Complex Gated Recurrent Neural Networks
- Discriminative Regularization for Generative Models
- Sub-1-us, Sub-20-nJ Pattern Classification in a Mixed-Signal Circuit Based on Embedded 180-nm Floating-Gate Memory Cell Arrays
- PCANet-II: When PCANet Meets the Second Order Pooling
- Research Frontiers in Transfer Learning -- a systematic and bibliometric review
- Phase autoencoder for rapid data-driven synchronization of rhythmic spatiotemporal patterns
- SD-GAN: Structural and Denoising GAN reveals facial parts under occlusion
- An Image Labeling Tool and Agricultural Dataset for Deep Learning
- Multi-Task Learning via Co-Attentive Sharing for Pedestrian Attribute Recognition
- Adaptive Multiscale Illumination-Invariant Feature Representation for Undersampled Face Recognition
- Deeply learning molecular structure-property relationships using attention- and gate-augmented graph convolutional network
- Automated proof synthesis for propositional logic with deep neural networks
- Efficient video annotation with visual interpolation and frame selection guidance
- Dynamic Relevance Learning for Few-Shot Object Detection
- Temporal Segment Networks for Action Recognition in Videos
- Hierarchical Deep Feature Fusion and Ensemble Learning for Enhanced Brain Tumor MRI Classification
- Feature Complementation Architecture for Visual Place Recognition
- DejaVid: Encoder-Agnostic Learned Temporal Matching for Video Classification
- An approximate Riemann solver approach in Physics-Informed Neural Networks for hyperbolic conservation laws
- Structural Design of Convolutional Neural Networks for Steganalysis
- 25 years of CNNs: Can we compare to human abstraction capabilities?
- Adversarial Attacks for Multi-view Deep Models
- Chart-Text: A Fully Automated Chart Image Descriptor
- Machine Unlearning for Robust DNNs: Attribution-Guided Partitioning and Neuron Pruning in Noisy Environments
- TruncQuant: Truncation-Ready Quantization for DNNs with Flexible Weight Bit Precision
- An Empirical Study on Leveraging Scene Graphs for Visual Question Answering
- The Cambrian Explosion of Mixed-Precision Matrix Multiplication for Quantized Deep Learning Inference
- cuDNN: Efficient Primitives for Deep Learning
- Evaluating Sensitivity Parameters in Smartphone-Based Gaze Estimation: A Comparative Study of Appearance-Based and Infrared Eye Trackers
- Generating 3D Molecular Structures Conditional on a Receptor Binding Site with Deep Generative Models
- Learning Models for Query by Vocal Percussion: A Comparative Study
- Structured Compression by Weight Encryption for Unstructured Pruning and Quantization
- MetaIQA: Deep Meta-learning for No-Reference Image Quality Assessment
- Designing and Training of A Dual CNN for Image Denoising
- GuidedMix-Net: Learning to Improve Pseudo Masks Using Labeled Images as Reference
- Topology-Aware Virtualization over Inter-Core Connected Neural Processing Units
- SecONNds: Secure Outsourced Neural Network Inference on ImageNet
- A Neural Rejection System Against Universal Adversarial Perturbations in Radio Signal Classification
- Big brothering the economy: nowcasting and forecasting with port satellite images
- Devanagari Digit Recognition using Quantum Machine Learning
- A Layered Self-Supervised Knowledge Distillation Framework for Efficient Multimodal Learning on the Edge
- Crossterm-Free Time-Frequency Representation Exploiting Deep Convolutional Neural Network
- EAR-U-Net: EfficientNet and attention-based residual U-Net for automatic liver segmentation in CT
- Pan-Sharpening with Color-Aware Perceptual Loss and Guided Re-Colorization
- ViLLa: A Neuro-Symbolic approach for Animal Monitoring
- Omni-sourced Webly-supervised Learning for Video Recognition
- Noise Modulation: Let Your Model Interpret Itself
- The Efficiency Misnomer
- Defensive Adversarial CAPTCHA: A Semantics-Driven Framework for Natural Adversarial Example Generation
- Generalizing Pooling Functions in Convolutional Neural Networks: Mixed, Gated, and Tree
- Metric Classification Network in Actual Face Recognition Scene
- Improving Visual Recognition using Ambient Sound for Supervision
- Video Frame Interpolation Transformer
- Deep Learning Theory Review: An Optimal Control and Dynamical Systems Perspective
- Binarizing MobileNet via Evolution-based Searching
- A Crack in the Bark: Leveraging Public Knowledge to Remove Tree-Ring Watermarks
- Semi-Tensor-Product Based Convolutional Neural Networks
- Stagewise Enlargement of Batch Size for SGD-based Learning
- The perceptual boost of visual attention is task-dependent in naturalistic settings
- Interior-Point Vanishing Problem in Semidefinite Relaxations for Neural Network Verification
- Machine learning approach to chance-constrained problems: An algorithm based on the stochastic gradient descent
- Optical Neural Networks
- Co-mining: Self-Supervised Learning for Sparsely Annotated Object Detection
- Multi-objective Neural Architecture Search with Almost No Training
- Thief, Beware of What Get You There: Towards Understanding Model Extraction Attack
- California Crop Yield Benchmark: Combining Satellite Image, Climate, Evapotranspiration, and Soil Data Layers for County-Level Yield Forecasting of Over 70 Crops
- Optimizing Genetic Algorithms with Multilayer Perceptron Networks for Enhancing TinyFace Recognition
- Bridging the Performance Gap between FGSM and PGD Adversarial Training
- Searching for TrioNet: Combining Convolution with Local and Global Self-Attention
- Rethinking of AlphaStar
- GLD-Road:A global-local decoding road network extraction model for remote sensing images
- HQFNN: A Compact Quantum-Fuzzy Neural Network for Accurate Image Classification
- Exploiting the Inherent Limitation of L0 Adversarial Examples
- Exploring Image Transforms derived from Eye Gaze Variables for Progressive Autism Diagnosis
- A Fast and Lightweight Model for Causal Audio-Visual Speech Separation
- Pareto-Frontier-aware Neural Architecture Generation for Diverse Budgets
- Exposing Semantic Segmentation Failures via Maximum Discrepancy Competition
- Efficient Winograd Convolution via Integer Arithmetic
- All SMILES Variational Autoencoder
- Aesthetics Without Semantics
- Learning to Linearize Under Uncertainty
- Confidential Inference via Ternary Model Partitioning
- Image retrieval method based on CNN and dimension reduction
- Dynamic Filtering with Large Sampling Field for ConvNets
- Forced Variational Integrator Networks for Prediction and Control of Mechanical Systems
- Improved Deep Spectral Convolution Network For Hyperspectral Unmixing With Multinomial Mixture Kernel and Endmember Uncertainty
- Enhancing decision support in crop production: Analyzing conformal prediction for uncertainty quantification
- Towards a Visual Turing Challenge
- A One-step Pruning-recovery Framework for Acceleration of Convolutional Neural Networks
- Learning Preference-Based Similarities from Face Images using Siamese Multi-Task CNNs
- Diagnosis of Pediatric Obstructive Sleep Apnea via Face Classification with Persistent Homology and Convolutional Neural Networks
- Partition Pruning: Parallelization-Aware Pruning for Deep Neural Networks
- HAVIR: HierArchical Vision to Image Reconstruction using CLIP-Guided Versatile Diffusion
- Graph Neural Networks for Natural Language Processing: A Survey
- ResPF: Residual Poisson Flow for Efficient and Physically Consistent Sparse-View CT Reconstruction
- Auto-DeepLab: Hierarchical Neural Architecture Search for Semantic Image Segmentation
- Dissecting Deep Neural Networks
- Neural Network for NILM Based on Operational State Change Classification
- Percival: Making In-Browser Perceptual Ad Blocking Practical With Deep Learning
- Where Should We Begin? A Low-Level Exploration of Weight Initialization Impact on Quantized Behaviour of Deep Neural Networks
- A feasibility study of deep neural networks for the recognition of banknotes regarding central bank requirements
- Extract and Merge: Merging extracted humans from different images utilizing Mask R-CNN
- Massively Parallel Methods for Deep Reinforcement Learning
- Model-guided Multi-path Knowledge Aggregation for Aerial Saliency Prediction
- Structured Sparsity Inducing Adaptive Optimizers for Deep Learning
- Multi-Point Proximity Encoding For Vector-Mode Geospatial Machine Learning
- Precipitation Forecasting via Multi-Scale Deconstructed ConvLSTM
- Towards LLM-Centric Multimodal Fusion: A Survey on Integration Strategies and Techniques
- Neural Architecture Search for Deep Face Recognition
- DAS-MAE: A self-supervised pre-training framework for universal and high-performance representation learning of distributed fiber-optic acoustic sensing
- Self-supervised One-Stage Learning for RF-based Multi-Person Pose Estimation
- Feature-Based Lie Group Transformer for Real-World Applications
- Reduced-Order Modeling through Machine Learning Approaches for Brittle Fracture Applications
- TIMING: Temporality-Aware Integrated Gradients for Time Series Explanation
- NIMO: a Nonlinear Interpretable MOdel
- Cheaper Pre-training Lunch: An Efficient Paradigm for Object Detection
- ConvTransformer: A Convolutional Transformer Network for Video Frame Synthesis
- Shield: Fast, Practical Defense and Vaccination for Deep Learning using JPEG Compression
- Compositional Convolutional Neural Networks: A Deep Architecture with Innate Robustness to Partial Occlusion
- HOTCAKE: Higher Order Tucker Articulated Kernels for Deeper CNN Compression
- Layered Motion Fusion: Lifting Motion Segmentation to 3D in Egocentric Videos
- StatsMerging: Statistics-Guided Model Merging via Task-Specific Teacher Distillation
- Learning to decompose for object detection and instance segmentation
- Bridging the Gap Between Neural Networks and Neuromorphic Hardware with A Neural Network Compiler
- SA-Net: Shuffle Attention for Deep Convolutional Neural Networks
- Regularization Matters: A Nonparametric Perspective on Overparametrized Neural Network
- MinCall - MinION end2end convolutional deep learning basecaller
- Imitating Deep Learning Dynamics via Locally Elastic Stochastic Differential Equations
- Person Re-Identification System at Semantic Level based on Pedestrian Attributes Ontology
- Triplet-based Deep Similarity Learning for Person Re-Identification
- Supervision-by-Registration: An Unsupervised Approach to Improve the Precision of Facial Landmark Detectors
- Video Deblurring with Deconvolution and Aggregation Networks
- Self-labelling via simultaneous clustering and representation learning
- Regression as Classification: Influence of Task Formulation on Neural Network Features
- Have you forgotten? A method to assess if machine learning models have forgotten data
- Hyb-KAN ViT: Hybrid Kolmogorov-Arnold Networks Augmented Vision Transformer
- Lifelong Robotic Reinforcement Learning by Retaining Experiences
- Multi-Branch Fully Convolutional Network for Face Detection
- Detecting Out-of-Distribution Inputs in Deep Neural Networks Using an Early-Layer Output
- Patchwork: A Patch-wise Attention Network for Efficient Object Detection and Segmentation in Video Streams
- Evolutionary Deep Learning to Identify Galaxies in the Zone of Avoidance
- Image Super-Resolution Using Deep Convolutional Networks
- On the loss landscape of a class of deep neural networks with no bad local valleys
- Artificial Intelligence : from Research to Application ; the Upper-Rhine Artificial Intelligence Symposium (UR-AI 2019)
- Multiple Sclerosis Lesion Segmentation -- A Survey of Supervised CNN-Based Methods
- LPM: Learnable Pooling Module for Efficient Full-Face Gaze Estimation
- The mutual exclusivity bias of bilingual visually grounded speech models
- Noise Optimization for Artificial Neural Networks
- One-Shot Image-to-Image Translation via Part-Global Learning with a Multi-adversarial Framework
- Key Points Estimation and Point Instance Segmentation Approach for Lane Detection
- Probabilistic Approach for Road-Users Detection
- Neither Quick Nor Proper -- Evaluation of QuickProp for Learning Deep Neural Networks
- Adaptive Latent Space Tuning for Non-Stationary Distributions
- Automatic Segmentation of Organs-at-Risk from Head-and-Neck CT using Separable Convolutional Neural Network with Hard-Region-Weighted Loss
- RAPDARTS: Resource-Aware Progressive Differentiable Architecture Search
- Fine-Grained Energy and Performance Profiling framework for Deep Convolutional Neural Networks
- The Effects of Image Pre- and Post-Processing, Wavelet Decomposition, and Local Binary Patterns on U-Nets for Skin Lesion Segmentation
- Style Normalization and Restitution for Domain Generalization and Adaptation
- DLFusion: An Auto-Tuning Compiler for Layer Fusion on Deep Neural Network Accelerator
- Moonshine: Distilling with Cheap Convolutions
- L3 Fusion: Fast Transformed Convolutions on CPUs
- Better Understanding Hierarchical Visual Relationship for Image Caption
- Writer-Aware CNN for Parsimonious HMM-Based Offline Handwritten Chinese Text Recognition
- Learning Discriminative Features Via Weights-biased Softmax Loss
- Vision Graph Prompting via Semantic Low-Rank Decomposition
- Temporal Unet: Sample Level Human Action Recognition using WiFi
- Transformation Driven Visual Reasoning
- Spirit Distillation: A Model Compression Method with Multi-domain Knowledge Transfer
- FPGA-Enabled Machine Learning Applications in Earth Observation: A Systematic Review
- PokeBNN: A Binary Pursuit of Lightweight Accuracy
- CMT: Convolutional Neural Networks Meet Vision Transformers
- A synthetic dataset for deep learning
- TATi-Thermodynamic Analytics ToolkIt: TensorFlow-based software for posterior sampling in machine learning applications
- Exploring Auxiliary Context: Discrete Semantic Transfer Hashing for Scalable Image Retrieval
- Paraphrase Augmented Task-Oriented Dialog Generation
- A Review of Computer Vision Methods in Network Security
- Research on deep learning in apple leaf disease recognition
- Pelican: A Deep Residual Network for Network Intrusion Detection
- Asymmetric CNN for image super-resolution
- X2Teeth: 3D Teeth Reconstruction from a Single Panoramic Radiograph
- A little goes a long way: Improving toxic language classification despite data scarcity
- Data augmentation instead of explicit regularization
- A Fast and Precise Method for Large-Scale Land-Use Mapping Based on Deep Learning
- GOLD-NAS: Gradual, One-Level, Differentiable
- Towards Generalizable Surgical Activity Recognition Using Spatial Temporal Graph Convolutional Networks
- Unsupervised Deep Tracking
- Semi-supervised Complex-valued GAN for Polarimetric SAR Image Classification
- Stein Variational Inference for Discrete Distributions
- Balancing Accuracy, Calibration, and Efficiency in Active Learning with Vision Transformers Under Label Noise
- What you need to know about the state-of-the-art computational models of object-vision: A tour through the models
- SocialAI: Benchmarking Socio-Cognitive Abilities in Deep Reinforcement Learning Agents
- Dynamic Neural Networks: A Survey
- Data Augmentation via Dependency Tree Morphing for Low-Resource Languages
- The History Began from AlexNet: A Comprehensive Survey on Deep Learning Approaches
- A Tutorial on Discriminative Clustering and Mutual Information
- Unrestricted Adversarial Examples via Semantic Manipulation
- Stabilizing Generative Adversarial Networks: A Survey
- Data-Driven Clustering via Parameterized Lloyd's Families
- Recent Advances in Deep Learning for Object Detection
- Video Saliency Prediction Using Enhanced Spatiotemporal Alignment Network
- LPRNet: Lightweight Deep Network by Low-rank Pointwise Residual Convolution
- Revisiting the Characteristics of Stochastic Gradient Noise and Dynamics
- Learning Deep Multi-Level Similarity for Thermal Infrared Object Tracking
- Interleaved Composite Quantization for High-Dimensional Similarity Search
- Rate Distortion For Model Compression: From Theory To Practice
- Data-dependent Initializations of Convolutional Neural Networks
- A mixed signal architecture for convolutional neural networks
- IC Networks: Remodeling the Basic Unit for Convolutional Neural Networks
- Adversarially Robust Neural Architectures
- Superpixel Segmentation via Convolutional Neural Networks with Regularized Information Maximization
- Teacher-Explorer-Student Learning: A Novel Learning Method for Open Set Recognition
- Deep Metric Learning Model for Imbalanced Fault Diagnosis
- The Algonauts Project 2021 Challenge: How the Human Brain Makes Sense of a World in Motion
- Multi-Cell Multi-Task Convolutional Neural Networks for Diabetic Retinopathy Grading
- Combining analysis of multi-parametric MR images into a convolutional neural network: Precise target delineation for vestibular schwannoma treatment planning
- Optimization Landscapes of Wide Deep Neural Networks Are Benign
- Unsupervised Part Mining for Fine-grained Image Classification
- Parametric Complexity Bounds for Approximating PDEs with Neural Networks
- Sequence Generation using Deep Recurrent Networks and Embeddings: A study case in music
- A Labeling-Free Approach to Supervising Deep Neural Networks for Retinal Blood Vessel Segmentation
- Deep Fragment Embeddings for Bidirectional Image Sentence Mapping
- Optimization Methods for Convolutional Sparse Coding
- Fast Fourier Transform-Based Spectral and Temporal Gradient Filtering for Differential Privacy
- Multi-Modal Hybrid Deep Neural Network for Speech Enhancement
- End-to-End Learning for the Deep Multivariate Probit Model
- VideoMCC: a New Benchmark for Video Comprehension
- EEG-based Drowsiness Estimation for Driving Safety using Deep Q-Learning
- Pipelining Split Learning in Multi-hop Edge Networks
- tempoGAN
- Butterfly-Net2: Simplified Butterfly-Net and Fourier Transform Initialization
- EmbedMask: Embedding Coupling for One-stage Instance Segmentation
- Category Anchor-Guided Unsupervised Domain Adaptation for Semantic Segmentation
- Conditionally Deep Hybrid Neural Networks Across Edge and Cloud
- PointHop++: A Lightweight Learning Model on Point Sets for 3D Classification
- Recurrent Deep Divergence-based Clustering for simultaneous feature learning and clustering of variable length time series
- Localized Compression: Applying Convolutional Neural Networks to Compressed Images
- A survey on Kornia: an Open Source Differentiable Computer Vision Library for PyTorch
- Domain Adaptation Gaze Estimation by Embedding with Prediction Consistency
- Semantic Image Segmentation with Deep Convolutional Nets and Fully Connected CRFs
- A Universal Music Translation Network
- Semi-supervised Grasp Detection by Representation Learning in a Vector Quantized Latent Space
- Some Improvements on Deep Convolutional Neural Network Based Image Classification
- MARS: Memory Attention-Aware Recommender System
- Three-dimensional convolutional neural network (3D-CNN) for heterogeneous material homogenization
- Attention-Guided Curriculum Learning for Weakly Supervised Classification and Localization of Thoracic Diseases on Chest Radiographs
- PointAtrousGraph: Deep Hierarchical Encoder-Decoder with Point Atrous Convolution for Unorganized 3D Points
- Predictively Encoded Graph Convolutional Network for Noise-Robust Skeleton-based Action Recognition
- ASK: Adaptively Selecting Key Local Features for RGB-D Scene Recognition
- Towards the Memorization Effect of Neural Networks in Adversarial Training
- Model Agnostic Combination for Ensemble Learning
- Going Deeper with Lean Point Networks
- Non-asymptotic Excess Risk Bounds for Classification with Deep Convolutional Neural Networks
- Adversarial Examples in Modern Machine Learning: A Review
- Exascale Deep Learning for Scientific Inverse Problems
- Improving Transformation Invariance in Contrastive Representation Learning
- Automatic Face Understanding: Recognizing Families in Photos
- Asynchronous Parallel Stochastic Gradient for Nonconvex Optimization
- Incorporating Luminance, Depth and Color Information by a Fusion-based Network for Semantic Segmentation
- ContextNet: A Click-Through Rate Prediction Framework Using Contextual information to Refine Feature Embedding
- Improved Deep Hashing with Soft Pairwise Similarity for Multi-label Image Retrieval
- Learning Better Features for Face Detection with Feature Fusion and Segmentation Supervision
- Compact Bilinear Pooling
- Joint Image Filtering with Deep Convolutional Networks
- Strategy and Benchmark for Converting Deep Q-Networks to Event-Driven Spiking Neural Networks
- Deleting object selective units in a fully-connected layer of deep convolutional networks improves classification performance
- Attending Category Disentangled Global Context for Image Classification
- Crowdsourcing Evaluation of Saliency-based XAI Methods
- Full-attention based Neural Architecture Search using Context Auto-regression
- An Adaptive and Fast Convergent Approach to Differentially Private Deep Learning
- Semantic Diversity versus Visual Diversity in Visual Dictionaries
- Latent Guided Sampling for Combinatorial Optimization
- ROSA: Addressing text understanding challenges in photographs via ROtated SAmpling
- VLMs Can Aggregate Scattered Training Patches
- SemiOccam: A Robust Semi-Supervised Image Recognition Network Using Sparse Labels
- Lockout: Sparse Regularization of Neural Networks
- Privacy Prediction of Images Shared on Social Media Sites Using Deep Features
- Generalizing from a Few Examples: A Survey on Few-Shot Learning
- Parameter Estimation with Dense and Convolutional Neural Networks Applied to the FitzHugh-Nagumo ODE
- Spatio-temporal Attention Model for Tactile Texture Recognition
- Tripartite Weight-Space Ensemble for Few-Shot Class-Incremental Learning
- Can NetGAN be improved on short random walks?
- End-To-End Trainable Video Super-Resolution Based on a New Mechanism for Implicit Motion Estimation and Compensation
- GrateTile: Efficient Sparse Tensor Tiling for CNN Processing
- Improved Residual Networks for Image and Video Recognition
- A Deep Neuro-Fuzzy Network for Image Classification
- BridgeNet: A Hybrid, Physics-Informed Machine Learning Framework for Solving High-Dimensional Fokker-Planck Equations
- Learning from Noise: Enhancing DNNs for Event-Based Vision through Controlled Noise Injection
- Deep Learning on Graphs: A Survey
- Purifying Shampoo: Investigating Shampoo's Heuristics by Decomposing its Preconditioner
- Models of Heavy-Tailed Mechanistic Universality
- Deep Markov Random Field for Image Modeling
- Fundamental tenis : resep meraih kemenangan / Tony Mottram
- Class-Aware Domain Adaptation for Improving Adversarial Robustness
- The Tensor Track VII: From Quantum Gravity to Artificial Intelligence
- Genome Sequence Classification for Animal Diagnostics with Graph Representations and Deep Neural Networks
- Deeply Activated Salient Region for Instance Search
- Measuring the Algorithmic Efficiency of Neural Networks
- Automatic Diagnosis of Short-Duration 12-Lead ECG using a Deep Convolutional Network
- Recent Advancements in Self-Supervised Paradigms for Visual Feature Representation
- Interpreting and Boosting Dropout from a Game-Theoretic View
- Self-supervised Representation Learning for Evolutionary Neural Architecture Search
- RAPIDNN: In-Memory Deep Neural Network Acceleration Framework
- Cogradient Descent for Dependable Learning
- Unifying Data, Model and Hybrid Parallelism in Deep Learning via Tensor Tiling
- HIEGNet: A Heterogenous Graph Neural Network Including the Immune Environment in Glomeruli Classification
- Neural Network Pruning with Residual-Connections and Limited-Data
- Reflective Decoding Network for Image Captioning
- A Deep Look into Neural Ranking Models for Information Retrieval
- A Survey of Deep Learning Video Super-Resolution
- RoadFormer : Local-Global Feature Fusion for Road Surface Classification in Autonomous Driving
- LapTool-Net: A Contextual Detector of Surgical Tools in Laparoscopic Videos Based on Recurrent Convolutional Neural Networks
- HierTrain: Fast Hierarchical Edge AI Learning with Hybrid Parallelism in Mobile-Edge-Cloud Computing
- Towards Omni-Supervised Face Alignment for Large Scale Unlabeled Videos
- Automated Segmentation for Hyperdense Middle Cerebral Artery Sign of Acute Ischemic Stroke on Non-Contrast CT Images
- Fine-grained Optimization of Deep Neural Networks
- Classification of Hoyle State Decay Branches in Active Target Time Projection Chamber using Neural Network
- ConMamba: Contrastive Vision Mamba for Plant Disease Detection
- Weak Novel Categories without Tears: A Survey on Weak-Shot Learning
- Auto-Labeling Data for Object Detection
- Dialog State Tracking with Reinforced Data Augmentation
- Random Registers for Cross-Domain Few-Shot Learning
- IF-TTN: Information Fused Temporal Transformation Network for Video Action Recognition
- Stacked Deconvolutional Network for Semantic Segmentation
- Pan-Arctic Permafrost Landform and Human-built Infrastructure Feature Detection with Vision Transformers and Location Embeddings
- Fault Localisation and Repair for DL Systems: An Empirical Study with LLMs
- MatConvNet - Convolutional Neural Networks for MATLAB
- Towards Recognizing New Semantic Concepts in New Visual Domains
- Rodrigues Network for Learning Robot Actions
- Tensor Normal Training for Deep Learning Models
- Through a Steerable Lens: Magnifying Neural Network Interpretability via Phase-Based Extrapolation
- Billion-scale semi-supervised learning for image classification
- Data-Efficient Learning of Feedback Policies from Image Pixels using Deep Dynamical Models
- pCAMP: Performance Comparison of Machine Learning Packages on the Edges
- Binarized Neural Architecture Search for Efficient Object Recognition
- HyPar: Towards Hybrid Parallelism for Deep Learning Accelerator Array
- UnrealCV: Connecting Computer Vision to Unreal Engine
- FlexiSAGA: A Flexible Systolic Array GEMM Accelerator for Sparse and Dense Processing
- Balancing Beyond Discrete Categories: Continuous Demographic Labels for Fair Face Recognition
- Self-supervised Latent Space Optimization with Nebula Variational Coding
- CLIP-driven rain perception: Adaptive deraining with pattern-aware network routing and mask-guided cross-attention
- Subspace Networks: Scaling Decentralized Training with Communication-Efficient Model Parallelism
- Energy Considerations for Large Pretrained Neural Networks
- WeightAlign: Normalizing Activations by Weight Alignment
- Ptolemy: Architecture Support for Robust Deep Learning
- Large scale digital prostate pathology image analysis combining feature extraction and deep neural network
- Fully-Convolutional Intensive Feature Flow Neural Network for Text Recognition
- ChannelExplorer: Exploring Class Separability Through Activation Channel Visualization
- Improving the Reproducibility of Deep Learning Software: An Initial Investigation through a Case Study Analysis
- SpatialFlow: Bridging All Tasks for Panoptic Segmentation
- DistillHash: Unsupervised Deep Hashing by Distilling Data Pairs
- Learning to Segment Object Candidates
- Knowledge Distillation: A Survey
- A Fully Convolutional Neural Network for Cardiac Segmentation in Short-Axis MRI
- Quotient Network -- A Network Similar to ResNet but Learning Quotients
- Understanding Ancient Coin Images
- Improving Interpretability for Computer-aided Diagnosis tools on Whole Slide Imaging with Multiple Instance Learning and Gradient-based Explanations
- 3D Skeleton-Based Action Recognition: A Review
- FuseFormer: Fusing Fine-Grained Information in Transformers for Video Inpainting
- Significance of Data Augmentation for Improving Cleft Lip and Palate Speech Recognition
- DS-TTS: Zero-Shot Speaker Style Adaptation from Voice Clips via Dynamic Dual-Style Feature Modulation
- Powerpropagation: A sparsity inducing weight reparameterisation
- MSN: Efficient Online Mask Selection Network for Video Instance Segmentation
- Prototype Guided Federated Learning of Visual Feature Representations
- Unified Adversarial Invariance
- SatDreamer360: Multiview-Consistent Generation of Ground-Level Scenes from Satellite Imagery
- Monitoring Robustness and Individual Fairness
- Condition-Invariant Multi-View Place Recognition
- Recursive speech separation for unknown number of speakers
- From Anchor Generation to Distribution Alignment: Learning a Discriminative Embedding Space for Zero-Shot Recognition
- Semi-Implicit Back Propagation
- Two-stage generative adversarial networks for document image binarization with color noise and background removal
- Sparse Array Capon Beamformer Design Availing Deep Learning
- Learning Instance-wise Sparsity for Accelerating Deep Models
- Device-Circuit-Architecture Co-Exploration for Computing-in-Memory Neural Accelerators
- Resource-Efficient Neural Networks for Embedded Systems
- Rethinking Softmax with Cross-Entropy: Neural Network Classifier as Mutual Information Estimator
- Learning To Characterize Adversarial Subspaces
- Privado: Practical and Secure DNN Inference with Enclaves
- COGNATE: Acceleration of Sparse Tensor Programs on Emerging Hardware using Transfer Learning
- Hyperspectral Anomaly Change Detection Based on Auto-encoder
- Automatic tracing of mandibular canal pathways using deep learning
- TanhExp: A Smooth Activation Function with High Convergence Speed for Lightweight Neural Networks
- SSAP: Single-Shot Instance Segmentation With Affinity Pyramid
- Image classification in frequency domain with 2SReLU: a second harmonics superposition activation function
- Co-designed pre-capture privacy optics for computer vision
- Typical Machine Learning Datasets as Low-Depth Quantum Circuits
- CHIP: Chameleon Hash-based Irreversible Passport for Robust Deep Model Ownership Verification and Active Usage Control
- On the Lipschitz Continuity of Set Aggregation Functions and Neural Networks for Sets
- VAEER: Visual Attention-Inspired Emotion Elicitation Reasoning
- Differentiable Augmentation for Data-Efficient GAN Training
- Error-Corrected Margin-Based Deep Cross-Modal Hashing for Facial Image Retrieval
- Drop Dropout on Single-Epoch Language Model Pretraining
- State Estimation and Control of Dynamic Systems from High-Dimensional Image Data
- Learning Deep Embeddings with Histogram Loss
- Predicting population neural activity in the Algonauts challenge using end-to-end trained Siamese networks and group convolutions
- Rethinking Convolutional Features in Correlation Filter Based Tracking
- Survey on Reliable Deep Learning-Based Person Re-Identification Models: Are We There Yet?
- Physics-inspired Energy Transition Neural Network for Sequence Learning
- Learning to Filter: Siamese Relation Network for Robust Tracking
- Multicolumn Networks for Face Recognition
- Energy-Embedded Neural Solvers for One-Dimensional Quantum Systems
- 50 Years of Automated Face Recognition
- TCM-Ladder: A Benchmark for Multimodal Question Answering on Traditional Chinese Medicine
- AugMix: A Simple Data Processing Method to Improve Robustness and Uncertainty
- MangoLeafViT: Leveraging Lightweight Vision Transformer with Runtime Augmentation for Efficient Mango Leaf Disease Classification
- BIRD: Behavior Induction via Representation-structure Distillation
- Improving Deep Hyperspectral Image Classification Performance with Spectral Unmixing
- Descriptor Matching with Convolutional Neural Networks: a Comparison to SIFT
- A Regressive Convolution Neural network and Support Vector Regression Model for Electricity Consumption Forecasting
- Bounding Box-Guided Diffusion for Synthesizing Industrial Images and Segmentation Map
- A Reverse Causal Framework to Mitigate Spurious Correlations for Debiasing Scene Graph Generation
- On the Validity of Head Motion Patterns as Generalisable Depression Biomarkers
- ESNet: An Efficient Symmetric Network for Real-time Semantic Segmentation
- Generalized Batch Normalization: Towards Accelerating Deep Neural Networks
- Accurate and Energy-Efficient Classification with Spiking Random Neural Network: Corrected and Expanded Version
- Interaction-Aware Trajectory Prediction of Connected Vehicles using CNN-LSTM Networks
- The Nonlinearity Coefficient - A Practical Guide to Neural Architecture Design
- Deep Cross-modal Hashing via Margin-dynamic-softmax Loss
- Advancing Image Super-resolution Techniques in Remote Sensing: A Comprehensive Survey
- WTEFNet: Real-Time Low-Light Object Detection for Advanced Driver Assistance Systems
- The Case for Strong Scaling in Deep Learning: Training Large 3D CNNs with Hybrid Parallelism
- Zero-Shot Learning with Sparse Attribute Propagation
- Toward Optimal Run Racing: Application to Deep Learning Calibration
- Deep Learning Bandgaps of Topologically Doped Graphene
- Finding a Needle in a Haystack: Tiny Flying Object Detection in 4K Videos using a Joint Detection-and-Tracking Approach
- Composition based crystal materials symmetry prediction using machine learning with enhanced descriptors
- Voxel-level Siamese Representation Learning for Abdominal Multi-Organ Segmentation
- Hardware Synthesis of State-Space Equations; Application to FPGA Implementation of Shallow and Deep Neural Networks
- Differential Gated Self-Attention
- RLScheduler: An Automated HPC Batch Job Scheduler Using Reinforcement Learning
- Balancing Robustness and Sensitivity using Feature Contrastive Learning
- Using Database Rule for Weak Supervised Text-to-SQL Generation
- SafeNet: An Assistive Solution to Assess Incoming Threats for Premises
- Unshuffling Data for Improved Generalization
- End-to-end Compression Towards Machine Vision: Network Architecture Design and Optimization
- Improving Augmentation and Evaluation Schemes for Semantic Image Synthesis
- Deep Embedded K-Means Clustering
- STEERAGE: Synthesis of Neural Networks Using Architecture Search and Grow-and-Prune Methods
- EquiReg: Equivariance Regularized Diffusion for Inverse Problems
- O2NA: An Object-Oriented Non-Autoregressive Approach for Controllable Video Captioning
- CrossNAS: A Cross-Layer Neural Architecture Search Framework for PIM Systems
- EEG-Inception: An Accurate and Robust End-to-End Neural Network for EEG-based Motor Imagery Classification
- Deep Sign: Enabling Robust Statistical Continuous Sign Language Recognition via Hybrid CNN-HMMs
- Interpretable Scaling Behavior in Sparse Subnetwork Representations of Quantum States
- Towards Understanding the Transferability of Deep Representations
- Analogy Search Engine: Finding Analogies in Cross-Domain Research Papers
- End-to-End Learned Image Compression with Quantized Weights and Activations
- Macromolecule Classification Based on the Amino-acid Sequence
- Multi-Modal Pedestrian Detection with Large Misalignment Based on Modal-Wise Regression and Multi-Modal IoU
- Deep Reinforcement Learning with Label Embedding Reward for Supervised Image Hashing
- The Resurrection of the ReLU
- From WiscKey to Bourbon: A Learned Index for Log-Structured Merge Trees
- Deep Learning with Low Precision by Half-wave Gaussian Quantization
- Test-time augmentation improves efficiency in conformal prediction
- Initializing Perturbations in Multiple Directions for Fast Adversarial Training
- Structure-aware scale-adaptive networks for cancer segmentation in whole-slide images
- Turning a Blind Eye: Explicit Removal of Biases and Variation from Deep Neural Network Embeddings
- Adversarial Attacks on Deep Learning Based mmWave Beam Prediction in 5G and Beyond
- S2AFormer: Strip Self-Attention for Efficient Vision Transformer
- A Survey of Behavior Learning Applications in Robotics -- State of the Art and Perspectives
- Streaming Object Detection for 3-D Point Clouds
- Continual Learning Beyond Experience Rehearsal and Full Model Surrogates
- Domain-Specific Suppression for Adaptive Object Detection
- BigDataBench: A Scalable and Unified Big Data and AI Benchmark Suite
- Large Scale Indexing of Generic Medical Image Data using Unbiased Shallow Keypoints and Deep CNN Features
- Expressive yet Efficient Feature Expansion with Adaptive Cross-Hadamard Products
- MolDesigner: Interactive Design of Efficacious Drugs with Deep Learning
- Transformer-Unet: Raw Image Processing with Unet
- A Novel ANN Structure for Image Recognition
- End to End Binarized Neural Networks for Text Classification
- Mitigating Generation Shifts for Generalized Zero-Shot Learning
- Joint Item Recommendation and Attribute Inference: An Adaptive Graph Convolutional Network Approach
- End-to-end feature fusion siamese network for adaptive visual tracking
- Automated defect classification in sewer closed circuit television inspections using deep convolutional neural networks
- EgoCoder: Intelligent Program Synthesis with Hierarchical Sequential Neural Network Model
- Relevance-driven Input Dropout: an Explanation-guided Regularization Technique
- SRNet: Improving Generalization in 3D Human Pose Estimation with a Split-and-Recombine Approach
- Characterizing the Reynolds number dependence of the chaotic attractor in two-dimensional turbulence with dimension-minimizing autoencoders
- How to trust unlabeled data? Instance Credibility Inference for Few-Shot Learning
- VeniBot: Towards Autonomous Venipuncture with Automatic Puncture Area and Angle Regression from NIR Images
- Long-Short Temporal Contrastive Learning of Video Transformers
- Learning Rich Nearest Neighbor Representations from Self-supervised Ensembles
- Sharpness-Aware Minimization with Z-Score Gradient Filtering
- MetaSlot: Break Through the Fixed Number of Slots in Object-Centric Learning
- Look, Listen and Learn - A Multimodal LSTM for Speaker Identification
- SISC: End-to-end Interpretable Discovery Radiomics-Driven Lung Cancer Prediction via Stacked Interpretable Sequencing Cells
- Compositional Scene Understanding through Inverse Generative Modeling
- Identifying and Exploiting Structures for Reliable Deep Learning
- The Physics of Local Optimization in Complex Disordered Systems
- Improving the Perceptual Quality of 2D Animation Interpolation
- A Novel BiLevel Paradigm for Image-to-Image Translation
- Detect-to-Retrieve: Efficient Regional Aggregation for Image Search
- Moment kernels: a simple and scalable approach for equivariance to rotations and reflections in deep convolutional networks
- Copresheaf Topological Neural Networks: A Generalized Deep Learning Framework
- One-Time Soft Alignment Enables Resilient Learning without Weight Transport
- PHISH in MESH: Korean Adversarial Phonetic Substitution and Phonetic-Semantic Feature Integration Defense
- Empowering Vector Graphics with Consistently Arbitrary Viewing and View-dependent Visibility
- PointAugment: an Auto-Augmentation Framework for Point Cloud Classification
- Visual Product Graph: Bridging Visual Products And Composite Images For End-to-End Style Recommendations
- Focusing on Shadows for Predicting Heightmaps from Single Remotely Sensed RGB Images with Deep Learning
- Multiple Descent: Design Your Own Generalization Curve
- ISTD-GCN: Iterative Spatial-Temporal Diffusion Graph Convolutional Network for Traffic Speed Forecasting
- Rethinking Recurrent Neural Networks and Other Improvements for Image Classification
- Indirect Domain Shift for Single Image Dehazing
- GateNet: Gating-Enhanced Deep Network for Click-Through Rate Prediction
- Kernel Quantile Embeddings and Associated Probability Metrics
- Position, Padding and Predictions: A Deeper Look at Position Information in CNNs
- Large-Scale Attribute-Object Compositions
- Learning to detect dysarthria from raw speech
- Mosaic: Data-Free Knowledge Distillation via Mixture-of-Experts for Heterogeneous Distributed Environments
- Dual Encoder Fusion U-Net (DEFU-Net) for Cross-manufacturer Chest X-ray Segmentation
- Learning and Interpreting Gravitational-Wave Features from CNNs with a Random Forest Approach
- Video-aided Unsupervised Grammar Induction
- Exploring the Possibility of TypiClust for Low-Budget Federated Active Learning
- ErpGS: Equirectangular Image Rendering enhanced with 3D Gaussian Regularization
- Conformal retrofitting via Riemannian manifolds: distilling task-specific graphs into pretrained embeddings
- Be Your Own Best Competitor! Multi-Branched Adversarial Knowledge Transfer
- Rethinking Bottleneck Structure for Efficient Mobile Network Design
- EfficientSeg: An Efficient Semantic Segmentation Network
- Detecting Volcano Deformation in InSAR using Deep learning
- Left ventricle segmentation By modelling uncertainty in prediction of deep convolutional neural networks and adaptive thresholding inference
- From What to How: Attributing CLIP's Latent Components Reveals Unexpected Semantic Reliance
- A Novel Convolutional Neural Network-Based Framework for Complex Multiclass Brassica Seed Classification
- Distributed Learning and its Application for Time-Series Prediction
- Predicting Onflow Parameters Using Transfer Learning for Domain and Task Adaptation
- Sample Efficient Adaptive Text-to-Speech
- Deep Learning and Bayesian Deep Learning Based Gender Prediction in Multi-Scale Brain Functional Connectivity
- Adversarial collision attacks on image hashing functions
- Class Incremental Online Streaming Learning
- Min-Entropy Latent Model for Weakly Supervised Object Detection
- Deep Learning Based Text Classification: A Comprehensive Review
- Self-Supervised Out-of-Distribution Detection in Brain CT Scans
- Learning a Domain Classifier Bank for Unsupervised Adaptive Object Detection
- OB3D: A New Dataset for Benchmarking Omnidirectional 3D Reconstruction Using Blender
- A Regularization-Guided Equivariant Approach for Image Restoration
- Shape-Texture Debiased Neural Network Training
- Fracking Deep Convolutional Image Descriptors
- Boosting Convolutional Features for Robust Object Proposals
- MomentsNet: a simple learning-free method for binary image recognition
- A Computational Method for Evaluating UI Patterns
- Recursive Inference for Variational Autoencoders
- Towards A Multi-agent System for Online Hate Speech Detection
- Coast Sargassum Level Estimation from Smartphone Pictures
- Jodi: Unification of Visual Generation and Understanding via Joint Modeling
- Co-AttenDWG: Co-Attentive Dimension-Wise Gating and Expert Fusion for Multi-Modal Offensive Content Detection
- An Abstraction Model for Semantic Segmentation Algorithms
- EMPNet: Neural Localisation and Mapping Using Embedded Memory Points
- Efficient SRAM-PIM Co-design by Joint Exploration of Value-Level and Bit-Level Sparsity
- SS-Auto: A Single-Shot, Automatic Structured Weight Pruning Framework of DNNs with Ultra-High Efficiency
- Lightweight Classification of IoT Malware based on Image Recognition
- DispVoxNets: Non-Rigid Point Set Alignment with Supervised Learning Proxies
- Do Large Language Models (Really) Need Statistical Foundations?
- Transaction-level Model Simulator for Communication-Limited Accelerators
- Greedy Network Enlarging
- Weight Evolution: Improving Deep Neural Networks Training through Evolving Inferior Weight Values
- AVD: Adversarial Video Distillation
- Engineering problems in machine learning systems
- Raiders of the Lost Architecture: Kernels for Bayesian Optimization in Conditional Parameter Spaces
- DNN-Chip Predictor: An Analytical Performance Predictor for DNN Accelerators with Various Dataflows and Hardware Architectures
- An Efficient Accelerator Design Methodology for Deformable Convolutional Networks
- AutoDrop: Training Deep Learning Models with Automatic Learning Rate Drop
- Path Aggregation Network for Instance Segmentation
- An Artificial Intelligence Model for Early Stage Breast Cancer Detection from Biopsy Images
- Guiding the Experts: Semantic Priors for Efficient and Focused MoE Routing
- Cell Detection in Microscopy Images with Deep Convolutional Neural Network and Compressed Sensing
- Performance and Generalizability Impacts of Incorporating Location Encoders into Deep Learning for Dynamic PM2.5 Estimation
- Generative Interventions for Causal Learning
- Incremental Meta-Learning via Indirect Discriminant Alignment
- Autocomp: A Powerful and Portable Code Optimizer for Tensor Accelerators
- Distinctive Feature Codec: An Adaptive Efficient Speech Representation for Depression Detection
- From Likelihood to Fitness: Improving Variant Effect Prediction in Protein and Genome Language Models
- Dense Human Body Correspondences Using Convolutional Networks
- Generative AI and foundation models in medical image
- Deep Learning-Based Detection of the Acute Respiratory Distress Syndrome: What Are the Models Learning?
- Data-driven forecasting of solar irradiance
- Self-Adaptive Partial Domain Adaptation
- Cross Attention Network for Semantic Segmentation
- Recover Canonical-View Faces in the Wild with Deep Neural Networks
- Image Resizing by Reconstruction from Deep Features
- LookUP: Vision-Only Real-Time Precise Underground Localisation for Autonomous Mining Vehicles
- LANCE: Efficient Low-Precision Quantized Winograd Convolution for Neural Networks Based on Graphics Processing Units
- Universal Adversarial Perturbations Generative Network for Speaker Recognition
- Feature Preserving Shrinkage on Bayesian Neural Networks via the R2D2 Prior
- TabSTAR: A Tabular Foundation Model for Tabular Data with Text Fields
- Beyond Discreteness: Finite-Sample Analysis of Straight-Through Estimator for Quantization
- A Caching Strategy Towards Maximal D2D Assisted Offloading Gain
- Clustering Effect of (Linearized) Adversarial Robust Models
- Graph Cut Segmentation Methods Revisited with a Quantum Algorithm
- Learning Global and Local Features of Normal Brain Anatomy for Unsupervised Abnormality Detection
- Less is more: Selecting informative and diverse subsets with balancing constraints
- 3D-OOCS: Learning Prostate Segmentation with Inductive Bias
- Joint Group Feature Selection and Discriminative Filter Learning for Robust Visual Object Tracking
- Difficulty Translation in Histopathology Images
- Evaluation Framework For Large-scale Federated Learning
- IDK Cascades: Fast Deep Learning by Learning not to Overthink
- PCONV: The Missing but Desirable Sparsity in DNN Weight Pruning for Real-time Execution on Mobile Devices
- Enabling Retrain-free Deep Neural Network Pruning using Surrogate Lagrangian Relaxation
- Promptable cancer segmentation using minimal expert-curated data
- Temporal Consistency Constrained Transferable Adversarial Attacks with Background Mixup for Action Recognition
- Learning to Discriminate Perturbations for Blocking Adversarial Attacks in Text Classification
- Structured Proxy Features for Multimodal NSCLC Survival Prediction from Pretreatment CT
- Semi-Supervised Self-Growing Generative Adversarial Networks for Image Recognition
- Radar Detection in the CBRS Band: Techniques, Challenges, and Future Directions
- A simple and effective postprocessing method for image classification
- A Deep Neural Network Surrogate Modeling Benchmark for Temperature Field Prediction of Heat Source Layout
- Centripetal SGD for Pruning Very Deep Convolutional Networks with Complicated Structure
- An Approach for Process Model Extraction By Multi-Grained Text Classification
- BGADAM: Boosting based Genetic-Evolutionary ADAM for Neural Network Optimization
- F1/10: An Open-Source Autonomous Cyber-Physical Platform
- AlignSeg: Feature-Aligned Segmentation Networks
- Hierarchical Representation Network for Steganalysis of QIM Steganography in Low-Bit-Rate Speech Signals
- Enhancing the Discriminative Feature Learning for Visible-Thermal Cross-Modality Person Re-Identification
- A Halo Merger Tree Generation and Evaluation Framework
- Correlated Logistic Model With Elastic Net Regularization for Multilabel Image Classification
- Machine Vision for Natural Gas Methane Emissions Detection Using an Infrared Camera
- EVM-Fusion: An Explainable Vision Mamba Architecture with Neural Algorithmic Fusion
- Behaviour Suite for Reinforcement Learning
- Towards Panoptic 3D Parsing for Single Image in the Wild
- Improving the Certified Robustness of Neural Networks via Consistency Regularization
- What Objective Does Self-paced Learning Indeed Optimize?
- Fast and Efficient Zero-Learning Image Fusion
- Machine Vision in the Context of Robotics: A Systematic Literature Review
- CycleMLP: A MLP-like Architecture for Dense Prediction
- Information Fusion in Attention Networks Using Adaptive and Multi-level Factorized Bilinear Pooling for Audio-visual Emotion Recognition
- Task-Aware Variational Adversarial Active Learning
- Learning Efficient Video Representation with Video Shuffle Networks
- A Systematic Review on Context-Aware Recommender Systems using Deep Learning and Embeddings
- DDR-ID: Dual Deep Reconstruction Networks Based Image Decomposition for Anomaly Detection
- Adversarial Learning of General Transformations for Data Augmentation
- The Local Elasticity of Neural Networks
- Natural vs Balanced Distribution in Deep Learning on Whole Slide Images for Cancer Detection
- Complexity and Diversity in Sparse Code Priors Improve Receptive Field Characterization of Macaque V1 Neurons
- Attention Transfer Network for Aspect-level Sentiment Classification
- Distilling Pixel-Wise Feature Similarities for Semantic Segmentation
- Learning Structured Twin-Incoherent Twin-Projective Latent Dictionary Pairs for Classification
- Seeing through Satellite Images at Street Views
- Critical Points of Random Neural Networks
- Rewrite the Stars
- Revealing Perceptible Backdoors, without the Training Set, via the Maximum Achievable Misclassification Fraction Statistic
- Making Convex Loss Functions Robust to Outliers using e-Exponentiated Transformation
- From Sound Representation to Model Robustness
- Deep localization of protein structures in fluorescence microscopy images
- Encrypted Speech Recognition using Deep Polynomial Networks
- Automated Performance Assessment in Transoesophageal Echocardiography with Convolutional Neural Networks
- Data augmentation with Symbolic-to-Real Image Translation GANs for Traffic Sign Recognition
- Multi-label classification of promotions in digital leaflets using textual and visual information
- A Two-Stage Data Selection Framework for Data-Efficient Model Training on Edge Devices
- SuperPure: Efficient Purification of Localized and Distributed Adversarial Patches via Super-Resolution GAN Models
- Entity Linking Meets Deep Learning: Techniques and Solutions
- Deep Learning-Driven Ultra-High-Definition Image Restoration: A Survey
- Small-to-Large Generalization: Data Influences Models Consistently Across Scale
- Data-Driven Breakthroughs and Future Directions in AI Infrastructure: A Comprehensive Review
- Do Convnets Learn Correspondence?
- Model Similarity Mitigates Test Set Overuse
- Forward and Backward Information Retention for Accurate Binary Neural Networks
- Deep Goal-Oriented Clustering
- BitSplit-Net: Multi-bit Deep Neural Network with Bitwise Activation Function
- Erased or Dormant? Rethinking Concept Erasure Through Reversibility
- Improved Trainable Calibration Method for Neural Networks on Medical Imaging Classification
- An Empirical Study of Graph Contrastive Learning
- SPGNet: Semantic Prediction Guidance for Scene Parsing
- Convolutional Hough Matching Networks
- An ETF view of Dropout regularization
- A New Unified Deep Learning Approach with Decomposition-Reconstruction-Ensemble Framework for Time Series Forecasting
- Fast Quantum Property Prediction via Deeper 2D and 3D Graph Networks
- Redox: Improving I/O Efficiency of Model Training Through File Redirection
- CryptoNN: Training Neural Networks over Encrypted Data
- COPYCAT: Practical Adversarial Attacks on Visualization-Based Malware Detection
- Learning Dynamic Alignment via Meta-filter for Few-shot Learning
- Enhancing Monte Carlo Dropout Performance for Uncertainty Quantification
- BosphorusSign22k Sign Language Recognition Dataset
- Representation Learning for Electronic Health Records
- Embodied Visual Recognition
- Neuromorphic Mimicry Attacks Exploiting Brain-Inspired Computing for Covert Cyber Intrusions
- Degree-Optimized Cumulative Polynomial Kolmogorov-Arnold Networks
- Contrastive Learning-Enhanced Trajectory Matching for Small-Scale Dataset Distillation
- Geometrically Regularized Transfer Learning with On-Manifold and Off-Manifold Perturbation
- From Pixels to Images: A Structural Survey of Deep Learning Paradigms in Remote Sensing Image Semantic Segmentation
- Deep Learning Enabled Segmentation, Classification and Risk Assessment of Cervical Cancer
- Deep Networks with Fast Retraining
- Road Network Metric Learning for Estimated Time of Arrival
- Improving Review Representations with User Attention and Product Attention for Sentiment Classification
- Synthesize then Compare: Detecting Failures and Anomalies for Semantic Segmentation
- Spectral Analysis of Latent Representations
- Guidelines for the Quality Assessment of Energy-Aware NAS Benchmarks
- Deep Stacked Hierarchical Multi-patch Network for Image Deblurring
- GenFT: A Generative Parameter-Efficient Fine-Tuning Method for Pretrained Foundation Models
- Accelerating Deep Unsupervised Domain Adaptation with Transfer Channel Pruning
- Body Part Regression for CT Images
- VET-DINO: Learning Anatomical Understanding Through Multi-View Distillation in Veterinary Imaging
- FDDWNet: A Lightweight Convolutional Neural Network for Real-time Sementic Segmentation
- TBT: Targeted Neural Network Attack with Bit Trojan
- Learning From Long-Tailed Data With Noisy Labels
- Can machines learn to see without visual databases?
- Dynamic Group Convolution for Accelerating Convolutional Neural Networks
- Training Deep Learning Based Denoisers without Ground Truth Data
- Multi-resolution Outlier Pooling for Sorghum Classification
- Spatiotemporal Attacks for Embodied Agents
- TransformerFusion: Monocular RGB Scene Reconstruction using Transformers
- Recognizing Instagram Filtered Images with Feature De-stylization
- Domain Adaptation for Multi-label Image Classification: a Discriminator-free Approach
- Minimizing Supervision in Multi-label Categorization
- Exploring Vision Transformers for Fine-grained Classification
- Bridging Predictive Coding and MDL: A Two-Part Code Framework for Deep Learning
- Vid2World: Crafting Video Diffusion Models to Interactive World Models
- FractalMamba++: Scaling Vision Mamba Across Resolutions via Hilbert Fractal Geometry
- Perceive Your Users in Depth: Learning Universal User Representations from Multiple E-commerce Tasks
- No-Reference Image Quality Assessment via Feature Fusion and Multi-Task Learning
- Generative Hierarchical Features from Synthesizing Images
- Investigating Emotion-Color Association in Deep Neural Networks
- Learning a Single Tucker Decomposition Network for Lossy Image Compression with Multiple Bits-Per-Pixel Rates
- Object Detection based on Region Decomposition and Assembly
- Tasks Structure Regularization in Multi-Task Learning for Improving Facial Attribute Prediction
- DOA estimation based on CNN for underwater acoustic array
- Self Distillation via Iterative Constructive Perturbations
- Enhancing Classification with Semi-Supervised Deep Learning Using Distance-Based Sample Weights
- Attention Based Pruning for Shift Networks
- Selective Structured State Space for Multispectral-fused Small Target Detection
- LQResNet: A Deep Neural Network Architecture for Learning Dynamic Processes
- Weisfeiler-Lehman Embedding for Molecular Graph Neural Networks
- Neural network-based modelling of unresolved stresses in a turbulent reacting flow with mean shear
- Boundary-Preserved Deep Denoising of the Stochastic Resonance Enhanced Multiphoton Images
- Anomaly Detection Based on Critical Paths for Deep Neural Networks
- Deep Ranking with Adaptive Margin Triplet Loss
- The Compressed Model of Residual CNDS
- MicroNets: Neural Network Architectures for Deploying TinyML Applications on Commodity Microcontrollers
- Understanding Task Representations in Neural Networks via Bayesian Ablation
- Bioinformatics and Medicine in the Era of Deep Learning
- Hard Class Rectification for Domain Adaptation
- Mitigating Uncertainty of Classifier for Unsupervised Domain Adaptation
- JUWELS Booster -- A Supercomputer for Large-Scale AI Research
- RepPoints: Point Set Representation for Object Detection
- StereoGAN: Bridging Synthetic-to-Real Domain Gap by Joint Optimization of Domain Translation and Stereo Matching
- Image Annotation based on Deep Hierarchical Context Networks
- On the geometry of generalization and memorization in deep neural networks
- DynaNoise: Dynamic Probabilistic Noise Injection for Defending Against Membership Inference Attacks
- WiPIN: Operation-free Passive Person Identification Using Wi-Fi Signals
- Agile Domain Adaptation
- LadderNet: Multi-path networks based on U-Net for medical image segmentation
- Data-driven Weight Initialization with Sylvester Solvers
- Few-Shot Object Detection with Attention-RPN and Multi-Relation Detector
- Universal Domain Adaptation through Self Supervision
- Convolutional Networks with Dense Connectivity
- Collaborative Teacher-Student Learning via Multiple Knowledge Transfer
- MetaAlign: Coordinating Domain Alignment and Classification for Unsupervised Domain Adaptation
- What Is Considered Complete for Visual Recognition?
- TBD: Benchmarking and Analyzing Deep Neural Network Training
- Memory-Free Generative Replay For Class-Incremental Learning
- An introduction to Neural Networks for Physicists
- Expert-Like Reparameterization of Heterogeneous Pyramid Receptive Fields in Efficient CNNs for Fair Medical Image Classification
- Brain tumor detection and multi‐classification using advanced deep learning techniques
- Performance Characterization of Distributed Deep Learning Strategies: A Quantitative Evaluation of DDP, FSDP, and Parameter Server Architectures on GPU Clusters
- Generating Socially Acceptable Perturbations for Efficient Evaluation of Autonomous Vehicles
- Recent Advances in Object Detection in the Age of Deep Convolutional Neural Networks
- High-dimensional structure underlying individual differences in naturalistic visual experience
- Advancing Generalization Across a Variety of Abstract Visual Reasoning Tasks
- Bandwidth-based Step-Sizes for Non-Convex Stochastic Optimization
- Additive Noise Annealing and Approximation Properties of Quantized Neural Networks
- Computer Vision Models Show Human-Like Sensitivity to Geometric and Topological Concepts
- Language Models That Walk the Talk: A Framework for Formal Fairness Certificates
- AGI-Elo: How Far Are We From Mastering A Task?
- VisDA-2021 Competition Universal Domain Adaptation to Improve Performance on Out-of-Distribution Data
- Full-stack Optimization for Accelerating CNNs with FPGA Validation
- Mamba-Adaptor: State Space Model Adaptor for Visual Recognition
- Parallel Layer Normalization for Universal Approximation
- Traceable Black-box Watermarks for Federated Learning
- SkyNet: A Champion Model for DAC-SDC on Low Power Object Detection
- DD-Ranking: Rethinking the Evaluation of Dataset Distillation
- Deep Priority Hashing
- Targeted Attack for Deep Hashing based Retrieval
- Deformation Robust Roto-Scale-Translation Equivariant CNNs
- MAT: A simple yet strong baseline for identifying self-admitted technical debt
- Coarse and fine-grained automatic cropping deep convolutional neural network
- Batch-Instance Normalization for Adaptively Style-Invariant Neural Networks
- FreqSelect: Frequency-Aware fMRI-to-Image Reconstruction
- ProMi: An Efficient Prototype-Mixture Baseline for Few-Shot Segmentation with Bounding-Box Annotations
- The Shallow End: Empowering Shallower Deep-Convolutional Networks through Auxiliary Outputs
- Energy-Aware Deep Learning on Resource-Constrained Hardware
- Overhead-MNIST: Machine Learning Baselines for Image Classification
- Synthetic History: Evaluating Visual Representations of the Past in Diffusion Models
- Sound Event Detection of Weakly Labelled Data with CNN-Transformer and Automatic Threshold Optimization
- Echo-Reconstruction: Audio-Augmented 3D Scene Reconstruction
- Learning Social Relation Traits from Face Images
- E2-Train: Training State-of-the-art CNNs with Over 80% Energy Savings
- PoLO: Proof-of-Learning and Proof-of-Ownership at Once with Chained Watermarking
- DragLoRA: Online Optimization of LoRA Adapters for Drag-based Image Editing in Diffusion Model
- CROSSBOW: Scaling Deep Learning with Small Batch Sizes on Multi-GPU Servers
- Build a Compact Binary Neural Network through Bit-level Sensitivity and Data Pruning
- Robust Federated Learning: The Case of Affine Distribution Shifts
- NeuroGen: Neural Network Parameter Generation via Large Language Models
- Spectral-Spatial Self-Supervised Learning for Few-Shot Hyperspectral Image Classification
- Improving Out-of-Domain Robustness with Targeted Augmentation in Frequency and Pixel Spaces
- GALA: Greedy ComputAtion for Linear Algebra in Privacy-Preserved Neural Networks
- Classifying Topological Charge in SU(3) Yang-Mills Theory with Machine Learning
- Urban land-use analysis using proximate sensing imagery: a survey
- Conditional Bures Metric for Domain Adaptation
- Waveform Phasicity Prediction from Arterial Sounds through Spectrogram Analysis using Convolutional Neural Networks for Limb Perfusion Assessment
- BPLF: A Bi-Parallel Linear Flow Model for Facial Expression Generation from Emotion Set Images
- Algorithm to Compilation Co-design: An Integrated View of Neural Network Sparsity
- Communication-Efficient Separable Neural Network for Distributed Inference on Edge Devices
- Responsible AI: Gender bias assessment in emotion recognition
- Global and Local Texture Randomization for Synthetic-to-Real Semantic Segmentation
- Understanding Intra-Class Knowledge Inside CNN
- DiCo: Revitalizing ConvNets for Scalable and Efficient Diffusion Modeling
- Neural network based limiter with transfer learning
- A deep learning method for image‐based subject‐specific local SAR assessment
- Transferable Contrastive Network for Generalized Zero-Shot Learning
- Deep learning-based synthetic-CT generation in radiotherapy and PET: a review
- Pix2seq: A Language Modeling Framework for Object Detection
- The NTNU System at the Interspeech 2020 Non-Native Children's Speech ASR Challenge
- Thalamocortical architectures for flexible cognition and efficient learning
- Towards a simplified model of primary visual cortex
- Towards a machine-learned Poisson solver for low-temperature plasma simulations in complex geometries
- GlitchNet: A Glitch Detection and Removal System for SEIS Records Based on Deep Learning
- Convolutional neuronal network for identifying single‐cell‐platelet–platelet‐aggregates in human whole blood using imaging flow cytometry
- Iterative Teaching by Label Synthesis
- Back to the Future: Joint Aware Temporal Deep Learning 3D Human Pose Estimation
- Self-supervised Learning of 3D Objects from Natural Images
- Learning Strict Identity Mappings in Deep Residual Networks
- DAF-NET: a saliency based weakly supervised method of dual attention fusion for fine-grained image classification
- A Generalized and Robust Method Towards Practical Gaze Estimation on Smart Phone
- Learning Local Shape Descriptors from Part Correspondences With Multi-view Convolutional Networks
- Unsupervised Domain Adaptation for Learning Eye Gaze from a Million Synthetic Images: An Adversarial Approach
- DPSeg: Dual-Prompt Cost Volume Learning for Open-Vocabulary Semantic Segmentation
- Cross-ethnicity Face Anti-spoofing Recognition Challenge: A Review
- MID-L: Matrix-Interpolated Dropout Layer with Layer-wise Neuron Selection
- A Multi-modal Fusion Network for Terrain Perception Based on Illumination Aware
- ℓp-Norm Multiple Kernel One-Class Fisher Null-Space
- Self-Supervised Policy Adaptation during Deployment
- Divergence Triangle for Joint Training of Generator Model, Energy-based Model, and Inference Model
- Differential equations as models of deep neural networks
- Semantically-Aware Attentive Neural Embeddings for Image-based Visual Localization
- Joint Graph Estimation and Signal Restoration for Robust Federated Learning
- Layer-specific spatiotemporal dynamics of feedforward and feedback in human visual object perception
- EgoDex: Learning Dexterous Manipulation from Large-Scale Egocentric Video
- A Survey of Label-noise Representation Learning: Past, Present and Future
- Convergence Analysis of the Last Iterate in Distributed Stochastic Gradient Descent with Momentum
- Self-supervised Visual Feature Learning with Deep Neural Networks: A Survey
- A Unified and Scalable Membership Inference Method for Visual Self-supervised Encoder via Part-aware Capability
- Joint Learning of Siamese CNNs and Temporally Constrained Metrics for Tracklet Association
- Modular Robot Control with Motor Primitives
- A 3D 2D convolutional Neural Network Model for Hyperspectral Image Classification
- Creating Robust Deep Neural Networks With Coded Distributed Computing for IoT Systems
- Learning Lyapunov Functions for Piecewise Affine Systems with Neural Network Controllers
- A Comparison for Anti-noise Robustness of Deep Learning Classification Methods on a Tiny Object Image Dataset: from Convolutional Neural Network to Visual Transformer and Performer
- MTVCrafter: 4D Motion Tokenization for Open-World Human Image Animation
- Quantum Computing and AI: Perspectives on Advanced Automation in Science and Engineering
- Recurrent Connectivity Aids Recognition of Partly Occluded Objects
- Mixing Real and Synthetic Data to Enhance Neural Network Training -- A Review of Current Approaches
- Continuous Homeostatic Reinforcement Learning for Self-Regulated Autonomous Agents
- ECG Heart-beat Classification Using Multimodal Image Fusion
- Learned Lightweight Smartphone ISP with Unpaired Data
- Reducing the Model Order of Deep Neural Networks Using Information Theory
- On the Reduction of Variance and Overestimation of Deep Q-Learning
- Learnable Histogram: Statistical Context Features for Deep Neural Networks
- An empirical study of task and feature correlations in the reuse of pre-trained models
- Unsupervised Domain Adaptation in Semantic Segmentation: a Review
- Discriminative Unsupervised Feature Learning with Exemplar Convolutional Neural Networks
- Rethinking Pretraining for Specialized Design Data: Evidence from the JONES-19 Cultural Design Dataset
- Analyze and Design Network Architectures by Recursion Formulas
- DAF:re: A Challenging, Crowd-Sourced, Large-Scale, Long-Tailed Dataset For Anime Character Recognition
- HWNet v2: An Efficient Word Image Representation for Handwritten Documents
- Tabular Benchmarks for Joint Architecture and Hyperparameter Optimization
- An In-Depth Analysis of Visual Tracking with Siamese Neural Networks
- Generating Adversarial Perturbation with Root Mean Square Gradient
- Universal Rules for Fooling Deep Neural Networks based Text Classification
- Model Selection with Nonlinear Embedding for Unsupervised Domain Adaptation
- Detecting Adversarial Examples from Sensitivity Inconsistency of Spatial-Transform Domain
- Pose Recognition with Cascade Transformers
- Deep Learning and Machine Vision for Food Processing: A Survey
- FONTNET: On-Device Font Understanding and Prediction Pipeline
- High-performance Semantic Segmentation Using Very Deep Fully Convolutional Networks
- DPPred: An Effective Prediction Framework with Concise Discriminative Patterns
- Multi-Task Driven Feature Models for Thermal Infrared Tracking
- Marigold: Affordable Adaptation of Diffusion-Based Image Generators for Image Analysis
- Efficient CNN Building Blocks for Encrypted Data
- Optimizing Millions of Hyperparameters by Implicit Differentiation
- Complex Wavelet SSIM based Image Data Augmentation
- I Am Going MAD: Maximum Discrepancy Competition for Comparing Classifiers Adaptively
- Adversarial Example in Remote Sensing Image Recognition
- Logographic Character Visual Pretraining via Semantic-based Contrastive Learning
- Multi-modality Latent Interaction Network for Visual Question Answering
- Scalable Modeling of Spatiotemporal Data using the Variational Autoencoder: an Application in Glaucoma
- Volcanic Clouds Detection through QCNN and Geostationary Satellite Multispectral Imagery
- Sequential Sentence Matching Network for Multi-turn Response Selection in Retrieval-based Chatbots
- From Artificial Neural Networks to Deep Learning for Music Generation -- History, Concepts and Trends
- Automated Model-Free Sorting of Single-Molecule Fluorescence Events Using a Deep Learning Based Hidden-State Model
- Transfer Learning for Material Classification using Convolutional Networks
- Generalization Error Bounds of Gradient Descent for Learning Over-parameterized Deep ReLU Networks
- AnonymousNet: Natural Face De-Identification with Measurable Privacy
- Boosting Adversarial Transferability through Enhanced Momentum
- Visual Image Reconstruction from Brain Activity via Latent Representation
- Deep Metric Learning by Online Soft Mining and Class-Aware Attention
- ViLT: Vision-and-Language Transformer Without Convolution or Region Supervision
- Deep Layer Aggregation
- Learning Feature-to-Feature Translator by Alternating Back-Propagation for Generative Zero-Shot Learning
- Cross-Database Micro-Expression Recognition: A Benchmark
- Attention-based Image Upsampling
- Super-fast rates of convergence for Neural Networks Classifiers under the Hard Margin Condition
- Few-shot Novel Category Discovery
- PCANet: An energy perspective
- Relational Teacher Student Learning with Neural Label Embedding for Device Adaptation in Acoustic Scene Classification
- Learning Soft Labels via Meta Learning
- DarwinML: A Graph-based Evolutionary Algorithm for Automated Machine Learning
- Improving Pairwise Ranking for Multi-label Image Classification
- Going Deeper Into Face Detection: A Survey
- DArFace: Deformation Aware Robustness for Low Quality Face Recognition
- Push for Quantization: Deep Fisher Hashing
- Learning mappings onto regularized latent spaces for biometric\n authentication
- Exposure: A White-Box Photo Post-Processing Framework
- Instance-aware Image Colorization with Controllable Textual Descriptions and Segmentation Masks
- TEAM: We Need More Powerful Adversarial Examples for DNNs
- Large Language Models for Computer-Aided Design: A Survey
- Putting It All into Context: Simplifying Agents with LCLMs
- Machine Learning on Volatile Instances
- Many-to-Many Voice Conversion using Cycle-Consistent Variational Autoencoder with Multiple Decoders
- Vanishing Nodes: Another Phenomenon That Makes Training Deep Neural Networks Difficult
- Detecting Colorized Images via Convolutional Neural Networks: Toward High Accuracy and Good Generalization
- Understanding and Improving Robustness of Vision Transformers through Patch-based Negative Augmentation
- Scene-Aware Error Modeling of LiDAR/Visual Odometry for Fusion-based Vehicle Localization
- Graph Neural Networks Exponentially Lose Expressive Power for Node Classification
- Survey of Visual-Semantic Embedding Methods for Zero-Shot Image Retrieval
- Video-based Person Re-Identification using Gated Convolutional Recurrent Neural Networks
- Automated Rib Fracture Detection of Postmortem Computed Tomography Images Using Machine Learning Techniques
- Feature Visualization in 3D Convolutional Neural Networks
- GradAug: A New Regularization Method for Deep Neural Networks
- Autonomous Robotic Pruning in Orchards and Vineyards: a Review
- Spherical Kernel for Efficient Graph Convolution on 3D Point Clouds
- Time-Varying Formation Controllers for Unmanned Aerial Vehicles Using Deep Reinforcement Learning
- Convolution Neural Network Architecture Learning for Remote Sensing Scene Classification
- Skin Lesion Classification Using Hybrid Deep Neural Networks
- An Online Deep Learning Approach Toward the Prediction of Power System Stresses Using Voltage Phasors
- RiFCN: Recurrent Network in Fully Convolutional Network for Semantic Segmentation of High Resolution Remote Sensing Images
- The Power of Contrast for Feature Learning: A Theoretical Analysis
- Enhancing Monocular Height Estimation via Sparse LiDAR-Guided Correction
- DP-TRAE: A Dual-Phase Merging Transferable Reversible Adversarial Example for Image Privacy Protection
- Equivariant and Invariant Reynolds Networks
- Bridging the Accuracy Gap for 2-bit Quantized Neural Networks (QNN)
- Efficient Contrastive Learning via Novel Data Augmentation and Curriculum Learning
- DeepID-Net: multi-stage and deformable deep convolutional neural networks for object detection
- Probabilistic Modeling for Novelty Detection with Applications to Fraud Identification
- NewsNet-SDF: Stochastic Discount Factor Estimation with Pretrained Language Model News Embeddings via Adversarial Networks
- Reproducing and Improving CheXNet: Deep Learning for Chest X-ray Disease Classification
- Impact of internal noise on convolutional neural networks
- AutoDNNchip: An Automated DNN Chip Predictor and Builder for Both FPGAs and ASICs
- A Survey on Data-Driven Modeling of Human Drivers' Lane-Changing Decisions
- Group Sparsity: The Hinge Between Filter Pruning and Decomposition for Network Compression
- OICSR: Out-In-Channel Sparsity Regularization for Compact Deep Neural Networks
- Deep Learning for Automatic Quality Grading of Mangoes: Methods and Insights
- Vision Transformers and Convolutional Neural Networks for Land Use Scene Classification
- Spintronic Neuromorphic Hardware Using Domain Wall-Based Neurons and Quantized Synapses
- Understanding Adversarial Examples from the Mutual Influence of Images and Perturbations
- 3D Correspondence Grouping with Compatibility Features
- NomMer: Nominate Synergistic Context in Vision Transformer for Visual Recognition
- The Application of Deep Learning for Lymph Node Segmentation: A Systematic Review
- Similarity-preserving Image-image Domain Adaptation for Person Re-identification
- SADet: Learning An Efficient and Accurate Pedestrian Detector
- Open Set Label Shift with Test Time Out-of-Distribution Reference
- Training Binary Neural Networks through Learning with Noisy Supervision
- A real-time hourly ozone prediction system using deep convolutional neural network
- Science Driven Innovations Powering Mobile Product: Cloud AI vs. Device AI Solutions on Smart Device
- Achieving 3D Attention via Triplet Squeeze and Excitation Block
- When are Deep Networks really better than Decision Forests at small sample sizes, and how?
- Generative Sensing: Transforming Unreliable Sensor Data for Reliable Recognition
- Deep Convolutional Ranking for Multilabel Image Annotation
- Black-Box Ripper: Copying black-box models using generative evolutionary algorithms
- Bombus Species Image Classification
- Fashion Retrieval via Graph Reasoning Networks on a Similarity Pyramid
- A Curriculum Domain Adaptation Approach to the Semantic Segmentation of Urban Scenes
- Info-Clustering: A Mathematical Theory for Data Clustering
- Generalisation in Neural Networks Does not Require Feature Overlap
- Neural Entropic Estimation: A faster path to mutual information estimation
- Anchors Based Method for Fingertips Position Estimation from a Monocular RGB Image using Deep Neural Network
- AugNet: End-to-End Unsupervised Visual Representation Learning with Image Augmentation
- Graph Neural Net using Analytical Graph Filters and Topology Optimization for Image Denoising
- Deep Learning based Cephalometric Landmark Identification using Landmark-dependent Multi-scale Patches
- Extreme 3D Face Reconstruction: Seeing Through Occlusions
- AutoShuffleNet: Learning Permutation Matrices via an Exact Lipschitz Continuous Penalty in Deep Convolutional Neural Networks
- Visual Affordance Prediction: Survey and Reproducibility
- Watermark retrieval from 3D printed objects via synthetic data training
- OmicsMapNet: Transforming omics data to take advantage of Deep Convolutional Neural Network for discovery
- Boosting Statistic Learning with Synthetic Data from Pretrained Large Models
- Analysis and Modeling of 3D Indoor Scenes
- An Evolution of CNN Object Classifiers on Low-Resolution Images
- D-CODA: Diffusion for Coordinated Dual-Arm Data Augmentation
- MasterSet: A Large-Scale Benchmark for Must-Cite Citation Recommendation in the AI/ML Literature
- Do Vision Transformers See Like Convolutional Neural Networks?
- Context Aware Machine Learning
- Real-time Federated Evolutionary Neural Architecture Search
- Hyperparameter Transfer Laws for Non-Recurrent Multi-Path Neural Networks
- Defining Benchmarks for Continual Few-Shot Learning
- Memory-efficient training with streaming dimensionality reduction
- Aligning Visual Prototypes with BERT Embeddings for Few-Shot Learning
- XtraLight-MedMamba for Classification of Neoplastic Tubular Adenomas
- Malicious URL Detection using Machine Learning: A Survey
- HandOcc: NeRF-based Hand Rendering with Occupancy Networks
- SpeechStew: Simply Mix All Available Speech Recognition Data to Train One Large Neural Network
- Multimodal Machine Learning: A Survey and Taxonomy
- NLNL: Negative Learning for Noisy Labels
- Theory-guided Data Science: A New Paradigm for Scientific Discovery from Data
- Data-Independent Structured Pruning of Neural Networks via Coresets
- OMNIA Faster R-CNN: Detection in the wild through dataset merging and soft distillation
- Context-Aware Group Captioning via Self-Attention and Contrastive Features
- Transfer-Learning-Aware Neuro-Evolution for Diseases Detection in Chest X-Ray Images
- Spatially Attentive Output Layer for Image Classification
- Kernel based regression with robust loss function via iteratively reweighted least squares
- Deep Learning: A Comprehensive Overview on Techniques, Taxonomy, Applications and Research Directions
- NINEPINS: Nuclei Instance Segmentation with Point Annotations
- Deep Joint Source Channel Coding for WirelessImage Transmission with OFDM
- OODTE: A Differential Testing Engine for the ONNX Optimizer
- ReMarNet: Conjoint Relation and Margin Learning for Small-Sample Image Classification
- Understanding the Mechanisms Behind Structural Influences on Link Prediction: A Case Study on FB15k-237
- Hacking Neural Networks: A Short Introduction
- Cross-Domain MLP and CNN Transfer Learning for Biological Signal Processing: EEG and EMG
- Focal-SAM: Focal Sharpness-Aware Minimization for Long-Tailed Classification
- End-to-end Flow Correlation Tracking with Spatial-temporal Attention
- TBNet:Pulmonary Tuberculosis Diagnosing System using Deep Neural Networks
- Farmland Parcel Delineation Using Spatio-temporal Convolutional Networks
- Role-Wise Data Augmentation for Knowledge Distillation
- Unsupervised Self-training Algorithm Based on Deep Learning for Optical Aerial Images Change Detection
- Semi-Supervised Noisy Student Pre-training on EfficientNet Architectures for Plant Pathology Classification
- AI Enabling Technologies: A Survey
- A Review on Object Pose Recovery: from 3D Bounding Box Detectors to Full 6D Pose Estimators
- Mini-batch graphs for robust image classification
- Leveraging Localization for Multi-camera Association
- Analysis of Deep Neural Networks with Quasi-optimal polynomial approximation rates
- DeepSymmetry : Using 3D convolutional networks for identification of tandem repeats and internal symmetries in protein structures
- Toxicity Prediction using Deep Learning
- LCP: A Low-Communication Parallelization Method for Fast Neural Network Inference in Image Recognition
- On the Generalization of Stochastic Gradient Descent with Momentum
- Orthogonal Over-Parameterized Training
- Relation-Shape Convolutional Neural Network for Point Cloud Analysis
- Cognitive Subscore Trajectory Prediction in Alzheimer's Disease
- Differentiated Backprojection Domain Deep Learning for Conebeam Artifact Removal
- What Deep CNNs Benefit from Global Covariance Pooling: An Optimization Perspective
- Few Sample Knowledge Distillation for Efficient Network Compression
- Approximation spaces of deep neural networks
- CostFilter-AD: Enhancing Anomaly Detection through Matching Cost Filtering
- Global Collinearity-aware Polygonizer for Polygonal Building Mapping in Remote Sensing
- Complexity-aware Adaptive Training and Inference for Edge-Cloud Distributed AI Systems
- Low-Precision Training of Large Language Models: Methods, Challenges, and Opportunities
- One Search Fits All: Pareto-Optimal Eco-Friendly Model Selection
- Towards the Resistance of Neural Network Watermarking to Fine-tuning
- Spatial Semantic Embedding Network: Fast 3D Instance Segmentation with Deep Metric Learning
- Multi-Scale Coarse-to-Fine Segmentation for Screening Pancreatic Ductal Adenocarcinoma
- Test time Adaptation through Perturbation Robustness
- Tracking Human-like Natural Motion Using Deep Recurrent Neural Networks
- Maximizing CNN Accelerator Efficiency Through Resource Partitioning
- Deep Learning Meets SAR
- A Novel Patch Convolutional Neural Network for View-based 3D Model Retrieval
- Weakly-Supervised Aspect-Based Sentiment Analysis via Joint Aspect-Sentiment Topic Embedding
- Finding and Visualizing Weaknesses of Deep Reinforcement Learning Agents
- Scale-aware Fast R-CNN for Pedestrian Detection
- Post-training Quantization with Multiple Points: Mixed Precision without Mixed Precision
- Towards Autonomous Driving of Personal Mobility with Small and Noisy Dataset using Tsallis-statistics-based Behavioral Cloning
- Multimodal Personal Ear Authentication Using Smartphones
- From images to insights: Using a convolutional neural network to improve powdery mildew severity detection in mungbean
- Comparative Study of Deep Learning Software Frameworks
- DFNets: Spectral CNNs for Graphs with Feedback-Looped Filters
- VIOLET : End-to-End Video-Language Transformers with Masked Visual-token Modeling
- Structured 2D Representation of 3D Data for Shape Processing
- FP-NAS: Fast Probabilistic Neural Architecture Search
- Towards Scalable Verification of Deep Reinforcement Learning
- End-to-end Video-level Representation Learning for Action Recognition
- Self-Supervised Learning by Estimating Twin Class Distributions
- Defending Adversarial Attacks by Correcting logits
- Leaderboard Incentives: Model Rankings under Strategic Post-Training
- An Energy-Efficient FPGA-based Deconvolutional Neural Networks Accelerator for Single Image Super-Resolution
- FPGA-Based CNN Inference Accelerator Synthesized from Multi-Threaded C Software
- Machine Learning Models that Remember Too Much
- A Framework for Generating Semantically Ambiguous Images to Probe Human and Machine Perception
- A Design Space for Live Music Agents
- Generating Images Instead of Retrieving Them
- Multi-Sample Dropout for Accelerated Training and Better Generalization
- Recursive Training of 2D-3D Convolutional Networks for Neuronal Boundary Detection
- Accelerated Learning with Robustness to Adversarial Regressors
- Fixed smooth convolutional layer for avoiding checkerboard artifacts in CNNs
- Democratizing AI: A Comparative Study in Deep Learning Efficiency and Future Trends in Computational Processing
- Proximal Alternating Direction Network: A Globally Converged Deep Unrolling Framework
- FOVI: A biologically-inspired foveated interface for deep vision models
- M2MRF: Many-to-Many Reassembly of Features for Tiny Lesion Segmentation in Fundus Images
- Predify: Augmenting deep neural networks with brain-inspired predictive coding dynamics
- Unlocking the potential of deep learning for marine ecology: overview, applications, and outlook
- Resource-Scalable CNN Synthesis for IoT Applications
- DeepAtrophy: Teaching a Neural Network to Differentiate Progressive Changes from Noise on Longitudinal MRI in Alzheimer's Disease
- The Deep Poincaré Map: A Novel Approach for Left Ventricle Segmentation
- Combining satellite imagery and machine learning to predict poverty
- Distorted Representation Space Characterization Through Backpropagated Gradients
- Image classification via a quantum-inspired strategy involving a mixture of experts
- General Purpose (GenP) Bioimage Ensemble of Handcrafted and Learned Features with Data Augmentation
- Semantics for Global and Local Interpretation of Deep Neural Networks
- A Possible Reason for why Data-Driven Beats Theory-Driven Computer Vision
- Real-time Multiple People Hand Localization in 4D Point Clouds
- Analytical aspects of non-differentiable neural networks
- Optimal Gradient Checkpoint Search for Arbitrary Computation Graphs
- Knowledge-Driven Machine Learning: Concept, Model and Case Study on Channel Estimation
- Convolutional Neural Networks for Classification of Alzheimer's Disease: Overview and Reproducible Evaluation
- Hybrid pattern recognition for charged particle tracking: Hough transform and convolutional neural efficiency networks
- Revisiting Locally Supervised Learning: an Alternative to End-to-end Training
- Learning to Transfer Graph Embeddings for Inductive Graph based Recommendation
- A Question Type Driven and Copy Loss Enhanced Frameworkfor Answer-Agnostic Neural Question Generation
- Inference of Recyclable Objects with Convolutional Neural Networks
- A Computing Kernel for Network Binarization on PyTorch
- Beyond Sharing Weights for Deep Domain Adaptation
- OCT Fingerprints: Resilience to Presentation Attacks
- Deep Video Super-Resolution Network Using Dynamic Upsampling Filters Without Explicit Motion Compensation
- The Neuroscience of Transformers
- The Multiverse Loss for Robust Transfer Learning
- Data Augmentation for Deep Learning-based Radio Modulation Classification
- The Automated Inspection of Opaque Liquid Vaccines
- Unsupervised Visual Representation Learning by Context Prediction
- Boosting Few-Shot Visual Learning With Self-Supervision
- Model predictive control design for dynamical systems learned by Long Short-Term Memory Networks
- Hardware Acceleration for Neural Networks: A Comprehensive Survey
- Pre-Training Estimators for Structural Models: Application to Consumer Search
- Contrastive Embedding for Generalized Zero-Shot Learning
- Hierarchical Severity Staging of Anterior Cruciate Ligament Injuries using Deep Learning with MRI Images
- Demystifying Parallel and Distributed Deep Learning
- Efficient Saliency Maps for Explainable AI
- Data Efficient Stagewise Knowledge Distillation
- Fed2
- 3D Convolution on RGB-D Point Clouds for Accurate Model-free Object Pose Estimation
- Dual Attention Suppression Attack: Generate Adversarial Camouflage in Physical World
- A Survey on Data Collection for Machine Learning: a Big Data -- AI Integration Perspective
- Deep learning using a biophysical model for Robust and Accelerated Reconstruction (RoAR) of quantitative and artifact-free R2* images
- Dense Pruning of Pointwise Convolutions in the Frequency Domain
- Neural population geometry: An approach for understanding biological and artificial neural networks
- Instance Segmentation Challenge Track Technical Report, VIPriors Workshop at ICCV 2021: Task-Specific Copy-Paste Data Augmentation Method for Instance Segmentation
- MERIT: Tensor Transform for Memory-Efficient Vision Processing on Parallel Architectures
- Multi-sense Definition Modeling using Word Sense Decompositions
- Making Graph Neural Networks Worth It for Low-Data Molecular Machine Learning
- Residual Encoder-Decoder Network for Deep Subspace Clustering
- A bibliometric retrospective of Computers & Security
- PRINS: Resistive CAM Processing in Storage
- Estimating and mapping Tokyo Window View Desirability using 3D model imagery and the hedonic price approach
- DeepFace: Closing the Gap to Human-Level Performance in Face Verification
- Deep Residual Bidir-LSTM for Human Activity Recognition Using Wearable Sensors
- Automated Steel Bar Counting and Center Localization with Convolutional Neural Networks
- Critique of Agent Model
- Identifying structural design principles shaping the computational abilities of recurrent neural networks
- Deep Context-Aware Kernel Networks
- MADAN: Multi-source Adversarial Domain Aggregation Network for Domain Adaptation
- Leveraging Uncertainty from Deep Learning for Trustworthy Materials Discovery Workflows
- Reinforced stochastic gradient descent for deep neural network learning
- From inference to prediction: how machine learning is reconfiguring science
- Learning Multi-view Deep Features for Small Object Retrieval in Surveillance Scenarios
- ExpandNets: Linear Over-parameterization to Train Compact Convolutional Networks
- Learning Phrase Representations using RNN Encoder–Decoder for Statistical Machine Translation
- Text Understanding from Scratch
- On the Orthogonality of Knowledge Distillation with Other Techniques: From an Ensemble Perspective
- Epistemic Neural Networks
- PoseConvGRU: A Monocular Approach for Visual Ego-motion Estimation by Learning
- Adversarial Cross-Domain Action Recognition with Co-Attention
- Lightweight and Robust Representation of Economic Scales from Satellite Imagery
- Podracer architectures for scalable Reinforcement Learning
- Blind Motion Deblurring with Cycle Generative Adversarial Networks
- Towards Defending Multiple ℓp-norm Bounded Adversarial Perturbations via Gated Batch Normalization
- Context-Aware Embeddings for Automatic Art Analysis
- Fast Fine-Grained Image Classification via Weakly Supervised Discriminative Localization
- Photovoltaic panel extraction from very high-resolution aerial imagery using region–line primitive association analysis and template matching
- PolyScientist: Automatic Loop Transformations Combined with Microkernels for Optimization of Deep Learning Primitives
- Unsupervised Lightweight Single Object Tracking with UHP-SOT++
- Robust Ensembling Network for Unsupervised Domain Adaptation
- Variational Resampling Based Assessment of Deep Neural Networks under Distribution Shift
- Jo-SRC: A Contrastive Approach for Combating Noisy Labels
- Reliable Shot Identification for Complex Event Detection via Visual-Semantic Embedding
- A Study for Universal Adversarial Attacks on Texture Recognition
- UnsuperPoint: End-to-end Unsupervised Interest Point Detector and Descriptor
- Pack-PTQ: Advancing Post-training Quantization of Neural Networks by Pack-wise Reconstruction
- Learning to Recognize Actionable Static Code Warnings (is Intrinsically Easy)
- Horizontal-to-Vertical Video Conversion
- DAmageNet: A Universal Adversarial Dataset
- Vision Mamba in Remote Sensing: A Comprehensive Survey of Techniques, Applications and Outlook
- Neural Networks Enhancement with Logical Knowledge
- Rethinking Generalisation
- Skeleton-Based Action Recognition with Synchronous Local and Non-local Spatio-temporal Learning and Frequency Attention
- Information based Deep Clustering: An experimental study
- Model Rubik's Cube: Twisting Resolution, Depth and Width for TinyNets
- Feature Alignment and Restoration for Domain Generalization and Adaptation
- Plumber: Diagnosing and Removing Performance Bottlenecks in Machine Learning Data Pipelines
- Convolutional Neural Networks using Logarithmic Data Representation
- A Dense Tensor Accelerator with Data Exchange Mesh for DNN and Vision Workloads
- Indicative Image Retrieval: Turning Blackbox Learning into Grey
- Large-Scale Generative Data-Free Distillation
- Exploring Semantic Segmentation on the DCT Representation
- Uncertainty-Aware Solar Flare Regression
- Malaria detection from RBC images using shallow Convolutional Neural Networks
- Gaussian Differential Privacy
- Unsupervised Feature Learning with K-means and An Ensemble of Deep Convolutional Neural Networks for Medical Image Classification
- Deep Neural Networks for Choice Analysis: A Statistical Learning Theory Perspective
- Estimating the Robustness of Classification Models by the Structure of the Learned Feature-Space
- Multi-Modal Music Information Retrieval: Augmenting Audio-Analysis with Visual Computing for Improved Music Video Analysis
- Non-Structured DNN Weight Pruning -- Is It Beneficial in Any Platform?
- Simulating multi-exit evacuation using deep reinforcement learning
- Machine learning: Trends, perspectives, and prospects
- A review of affective computing: From unimodal analysis to multimodal fusion
- Visual features for context-aware speech recognition
- Neural Networks with Inputs Based on Domain of Dependence and A Converging Sequence for Solving Conservation Laws, Part I: 1D Riemann Problems
- Vision-based Automated Bridge Component Recognition Integrated With High-level Scene Understanding
- Automated Vision-based Bridge Component Extraction Using Multiscale Convolutional Neural Networks
- Exploiting Operation Importance for Differentiable Neural Architecture Search
- What and Where: Modeling Skeletons from Semantic and Spatial Perspectives for Action Recognition
- Style-aware gloss control for generative non-photorealistic rendering
- BUZz: BUffer Zones for defending adversarial examples in image classification
- Straight to Shapes++: Real-time Instance Segmentation Made More Accurate
- Dialect Identification in Nuanced Arabic Tweets Using Farasa Segmentation and AraBERT
- STELA: A Real-Time Scene Text Detector with Learned Anchor
- DropSample: A New Training Method to Enhance Deep Convolutional Neural Networks for Large-Scale Unconstrained Handwritten Chinese Character Recognition
- AI-ready Snow Radar Echogram Dataset (SRED) for climate change monitoring
- Inducing Functions through Reinforcement Learning without Task Specification
- Hierarchical Recurrent Attention Networks for Structured Online Maps
- Deep reinforcement learning for the control of conjugate heat transfer with application to workpiece cooling
- Neural information retrieval: at the end of the early years
- Smoothed Gaussian Mixture Models for Video Classification and Recommendation
- Funnel Activation for Visual Recognition
- Sparse-GAN: Sparsity-constrained Generative Adversarial Network for Anomaly Detection in Retinal OCT Image
- Mapping, Localization and Path Planning for Image-based Navigation using Visual Features and Map
- Celo2: Towards Learned Optimization Free Lunch
- Modeling language and cognition with deep unsupervised learning: a tutorial overview
- SIMDRAM: An End-to-End Framework for Bit-Serial SIMD Computing in DRAM
- Performance of deep learning vs machine learning in plant leaf disease detection
- Efficient Model Performance Estimation via Feature Histories
- Learning Transformation Synchronization
- Portraying double Higgs at the Large Hadron Collider
- Evaluating data augmentation for financial time series classification
- AgriVariant: Variant Effect Prediction using DeepChem-Variant for Precision Breeding in Rice
- State of the Art on Neural Rendering
- Residue Number System-Based Solution for Reducing the Hardware Cost of a Convolutional Neural Network
- Stochastic encoding of graphs in deep learning allows for complex analysis of gender classification in resting-state and task functional brain networks from the UK Biobank
- Powder-Bed Fusion Process Monitoring by Machine Vision With Hybrid Convolutional Neural Networks
- GPU-Accelerated Genetic Programming for Symbolic Regression with Beagle Framework
- Analyzing and Improving the Image Quality of StyleGAN
- Vision Transformers in Precision Agriculture: A Comprehensive Survey
- QuantumToolbox.jl: An efficient Julia framework for simulating open quantum systems
- Scalable Multi-Task Learning for Particle Collision Event Reconstruction with Heterogeneous Graph Neural Networks
- Optimizing Mouse Dynamics for User Authentication by Machine Learning: Addressing Data Sufficiency, Accuracy-Practicality Trade-off, and Model Performance Challenges
- NGENT: Next-Generation AI Agents Must Integrate Multi-Domain Abilities to Achieve Artificial General Intelligence
- Crowded Scene Analysis: A Survey
- Automated identification of media bias in news articles: an interdisciplinary literature review
- MIRAGE: Robust multi-modal architectures translate fMRI-to-image models from vision to mental imagery
- Towards the Automatic Classification of Avian Flight Calls for Bioacoustic Monitoring
- Grounded Compositional Semantics for Finding and Describing Images with Sentences
- BoTShark: A Deep Learning Approach for Botnet Traffic Detection
- Mastering the game of Go with deep neural networks and tree search
- Classification of Android apps and malware using deep neural networks
- There’s plenty of room at the Top: What will drive computer performance after Moore’s law?
- Invariant object recognition is a personalized selection of invariant features in humans, not simply explained by hierarchical feed-forward vision models
- Convolutional Autoencoders for Data Compression and Anomaly Detection in Small Satellite Technologies
- SCOPE-MRI: Bankart Lesion Detection as a Case Study in Data Curation and Deep Learning for Challenging Diagnoses
- Exploring internal representation of self-supervised networks: few-shot learning abilities and comparison with human semantics and recognition of objects
- Leveraging Depth Maps and Attention Mechanisms for Enhanced Image Inpainting
- Do We Need More Training Data?
- Enhancing Vulnerability Reports with Automated and Augmented Description Summarization
- Frequency Feature Fusion Graph Network For Depression Diagnosis Via fNIRS
- Quantifying the Influence of Climate on Storm Activity Using Machine Learning
- Full interpretation of minimal images
- Enhancing event reconstruction for γ-ray particle detector arrays using transformers
- Evolution of Accuracy and Visual-Cognitive Errors in a Decade of Vision-Language AI Models
- You Don't Need Strong Assumptions: Visual Representation Learning via Temporal Differences
- DeepTox: Toxicity Prediction using Deep Learning
- Learning efficient haptic shape exploration with a rigid tactile sensor array
- Declinio Global de Especies
- Facial attractiveness of cleft patients: a direct comparison between artificial-intelligence-based scoring and conventional rater groups
- Hybrid Quantum-MambaVision: A Quantum-enhanced State Space Model for Calibrated Mixed-type Wafer Defect Detection
- Maximum Likelihood Reinforcement Learning
- Faster Predictive Coding Networks via Better Initialization
- Two-Stream Video Classification with Cross-Modality Attention
- Nested Learning: The Illusion of Deep Learning Architectures
- Data-Driven Stabilization of Unknown Linear-Threshold Network Dynamics
- Breast Cancer Detection from Multi-View Screening Mammograms with Visual Prompt Tuning
- AI Supply Chains: An Emerging Ecosystem of AI Actors, Products, and Services
- FusionNet: Multi-model Linear Fusion Framework for Low-light Image Enhancement
- Real-time Driver Drowsiness Detection for Android Application Using Deep Neural Networks Techniques
- Mitigating Bias in Facial Recognition Systems: Centroid Fairness Loss Optimization
- SoCodeCNN: Program Source Code for Visual CNN Classification Using Computer Vision Methodology
- DEEP-GAP: Deep-learning Evaluation of Execution Parallelism in GPU Architectural Performance
- AIBuildAI: An AI Agent for Automatically Building AI Models
- RoboVerse: Towards a Unified Platform, Dataset and Benchmark for Scalable and Generalizable Robot Learning
- PyViT-FUSE: A Foundation Model for Multi-Sensor Earth Observation Data
- 3DPyranet Features Fusion for Spatio-temporal Feature Learning
- Towards a deep learning approach for classifying treatment response in glioblastomas
- Experimental neuromorphic computing based on quantum memristor
- Gradient Descent as a Shrinkage Operator for Spectral Bias
- Human-level control through deep reinforcement learning
- Mastering the game of Go without human knowledge
- New types of deep neural network learning for speech recognition and related applications: an overview
- A Comparison of Dense Region Detectors for Image Search and Fine-Grained Classification
- From Image-Level to Pixel-Level Labeling with Convolutional Networks
- Parallel Training of Deep Networks with Local Updates
- Deep bilateral learning for real-time image enhancement
- Overcoming catastrophic forgetting in neural networks
- Video Classification with Channel-Separated Convolutional Networks
- Classification of SARS-CoV-2 Variants through The Epistatical Circos Plots with Convolutional Neural Networks
- Deep learning for time series classification
- Certifying Joint Adversarial Robustness for Model Ensembles
- Efficient Segmentation: Learning Downsampling Near Semantic Boundaries
- DIVE: Inverting Conditional Diffusion Models for Discriminative Tasks
- Game Theoretical Adversarial Deep Learning With Variational Adversaries
- Adversarial Deep Learning Models with Multiple Adversaries
- Person Re-Identification by Deep Joint Learning of Multi-Loss Classification
- Foveation-based Mechanisms Alleviate Adversarial Examples
- Ensemble deep learning: A review
- Contrastive Multiview Coding
- Revisiting Neural Retrieval on Accelerators
- When Celebrities Endorse Politicians: Analyzing the Behavior of Celebrity Followers in the 2016 U.S. Presidential Election
- Preprocessing Methods and Pipelines of Data Mining: An Overview
- VGG Fine-tuning for Cooking State Recognition
- KCP: Kernel Cluster Pruning for Dense Labeling Neural Networks
- Neuron Coverage-Guided Domain Generalization
- Channel-wise Dynamic Knowledge Distillation via Adaptive Sample Generation for Action Recognition
- Learning Temporally Invariant and Localizable Features via Data Augmentation for Video Recognition
- Towards artificial general intelligence with hybrid Tianjic chip architecture
- Self-supervised Learning with Fully Convolutional Networks
- Enactive Artificial Intelligence: A Decision-Centric Architecture for Complex Systems
- Stochastic Model Pruning via Weight Dropping Away and Back
- Selective Pseudo-Labeling with Reinforcement Learning for Semi-Supervised Domain Adaptation
- Revealing Fine Structures of the Retinal Receptive Field by Deep Learning Networks
- Acoustic Scene Classification Using Bilinear Pooling on Time-liked and Frequency-liked Convolution Neural Network
- A Relation-Augmented Fully Convolutional Network for Semantic Segmentation in Aerial Scenes
- Learning Object Scale With Click Supervision for Object Detection
- Train and Deploy an Image Classifier for Disaster Response
- Learn to Interpret Atari Agents
- A Multi-task Neural Approach for Emotion Attribution, Classification and Summarization
- HCNet: Hierarchical Context Network for Semantic Segmentation
- Learning Temporal Pose Estimation from Sparsely-Labeled Videos
- Image Classification base on PCA of Multi-view Deep Representation
- OUI Need to Talk About Weight Decay: A New Perspective on Overfitting Detection
- Interpretable BoW Networks for Adversarial Example Detection
- Data Movement Is All You Need: A Case Study on Optimizing Transformers
- On Integrating Information Visualization Techniques into Data Mining: A Review
- Low-Rank Matrix Approximation for Neural Network Compression
- Pulsar Candidate Identification with Artificial Intelligence Techniques
- Learning deep spatiotemporal features for video captioning
- Advances in Deep Learning for Hyperspectral Image Analysis—Addressing Challenges Arising in Practical Imaging Scenarios
- Eye state recognition based on deep integrated neural network and transfer learning
- A deep learning framework for neuroscience
- Geometry aware inference of steady state PDEs using Equivariant Neural Fields representations
- Teacher-Student Training for Robust Tacotron-based TTS
- Analyzing Representations inside Convolutional Neural Networks
- A Discriminatively Learned CNN Embedding for Person Reidentification
- Multi-Objective Convolutional Neural Networks for Robot Localisation and 3D Position Estimation in 2D Camera Images
- Design Choices That Matter: A Functional ANOVA Analysis for Remote Sensing Multi-Label Classification
- Learning Abstract Classes using Deep Learning
- Handwritten isolated Bangla compound character recognition: A new benchmark using a novel deep learning approach
- Deep Semantic Multimodal Hashing Network for Scalable Image-Text and Video-Text Retrievals
- HENet: Forcing a Network to Think More for Font Recognition
- Monocular human pose estimation: A survey of deep learning-based methods
- Perturbation analysis of gradient-based adversarial attacks
- Careful analysis of XRD patterns with Attention
- Comparison of augmentation and pre-processing for deep learning and chemometric classification of infrared spectra
- Two Novel Performance Improvements for Evolving CNN Topologies
- Music Artist Classification with Convolutional Recurrent Neural Networks
- Prolongation of SMAP to Spatiotemporally Seamless Coverage of Continental U.S. Using a Deep Learning Neural Network
- Visual Text Correction
- Nonparametric regression using deep neural networks with ReLU activation function
- Efficient Yet Deep Convolutional Neural Networks for Semantic Segmentation
- Occam’s Razor for Big Data? On Detecting Quality in Large Unstructured Datasets
- Transductive Zero-Shot Hashing for Multilabel Image Retrieval
- Visual Servoing for Pose Control of Soft Continuum Arm in a Structured Environment
- A deep learning approach for pose estimation from volumetric OCT data
- Deep multi-task learning for a geographically-regularized semantic segmentation of aerial images
- 3D Sketching using Multi-View Deep Volumetric Prediction
- FoodNet: Recognizing Foods Using Ensemble of Deep Networks
- Data augmentation in microscopic images for material data mining
- Deep convolutional neural networks for uncertainty propagation in random fields
- Text-Attentional Convolutional Neural Network for Scene Text Detection
- Video Imprint
- Exploring the Interchangeability of CNN Embedding Spaces
- Comprehensive Privacy Analysis of Deep Learning: Passive and Active White-box Inference Attacks against Centralized and Federated Learning
- Two-stream Fusion Model for Dynamic Hand Gesture Recognition using 3D-CNN and 2D-CNN Optical Flow guided Motion Template
- Bayesian Optimization for Iterative Learning
- Beware the Black-Box: On the Robustness of Recent Defenses to Adversarial Examples
- Recyclable Waste Identification Using CNN Image Recognition and Gaussian Clustering
- Rethinking multiscale cardiac electrophysiology with machine learning and predictive modelling
- An Improvement of Data Classification Using Random Multimodel Deep Learning (RMDL)
- Neural network models for the anisotropic Reynolds stress tensor in turbulent channel flow
- Representation Learning for Natural Language Processing
- Deep neural networks for record counting in historical handwritten documents
- Differentiable Neural Architecture Learning for Efficient Neural Network Design
- Efficient Inference via Universal LSH Kernel
- Grafted network for person re-identification
- MDSSD: Multi-scale Deconvolutional Single Shot Detector for Small Objects
- Bayesian multiscale deep generative model for the solution of high-dimensional inverse problems
- SegICP: Integrated deep semantic segmentation and pose estimation
- NaturalAE: Natural and robust physical adversarial examples for object detectors
- MobiVSR: A Visual Speech Recognition Solution for Mobile Devices
- Mind mappings: enabling efficient algorithm-accelerator mapping space search
- DeepIGeoS: A Deep Interactive Geodesic Framework for Medical Image Segmentation
- Land cover classification from multi-temporal, multi-spectral remotely sensed imagery using patch-based recurrent neural networks
- A deep convolutional neural network to analyze position averaged convergent beam electron diffraction patterns
- DAQN: Deep Auto-encoder and Q-Network
- Retinal vessel segmentation based on Fully Convolutional Neural Networks
- Marginal Contribution Feature Importance -- an Axiomatic Approach for The Natural Case
- SR-clustering: Semantic regularized clustering for egocentric photo streams segmentation
- Memory Based Online Learning of Deep Representations from Video Streams
- Do citations and readership identify seminal publications?
- Exploring Deep and Recurrent Architectures for Optimal Control
- Copycat CNN: Are random non-Labeled data enough to steal knowledge from black-box models?
- On Intrinsic Dataset Properties for Adversarial Machine Learning
- Q-CapsNets: A Specialized Framework for Quantizing Capsule Networks
- Good Features to Correlate for Visual Tracking
- Practical Convex Formulation of Robust One-hidden-layer Neural Network Training
- Neural Networks as Explicit Word-Based Rules
- Ranking to Learn and Learning to Rank: On the Role of Ranking in Pattern Recognition Applications
- EEG Classification by factoring in Sensor Configuration
- Symbiotic Attention with Privileged Information for Egocentric Action Recognition
- Feature Fusion Vision Transformer for Fine-Grained Visual Categorization
- Rapid Exact Signal Scanning With Deep Convolutional Neural Networks
- GEVO: GPU Code Optimization using Evolutionary Computation
- Deep Learning -- A first Meta-Survey of selected Reviews across Scientific Disciplines, their Commonalities, Challenges and Research Impact
- Photonic Convolution Neural Network Based on Interleaved Time-Wavelength Modulation
- The Role of Momentum Parameters in the Optimal Convergence of Adaptive Polyak's Heavy-ball Methods
- A survey of deep learning techniques for autonomous driving
- Vulnerability of Appearance-based Gaze Estimation
- Online Adaptation through Meta-Learning for Stereo Depth Estimation
- Systolic-CNN: An OpenCL-defined Scalable Run-time-flexible FPGA Accelerator Architecture for Accelerating Convolutional Neural Network Inference in Cloud/Edge Computing
- Identifying Lead Water Service Lines Using Ultrasonic Stress Wave Propagation and 1D-Convolutional Neural Network
- Unsupervised Deep Feature Extraction for Remote Sensing Image Classification
- Enhancing Privacy in Semantic Communication over Wiretap Channels leveraging Differential Privacy
- FRDet: Balanced and Lightweight Object Detector based on Fire-Residual Modules for Embedded Processor of Autonomous Driving
- OSVNet: Convolutional Siamese Network for Writer Independent Online\n Signature Verification
- Distilling Localization for Self-Supervised Representation Learning
- I-Con: A Unifying Framework for Representation Learning
- What Makes for a Good Saliency Map? Comparing Strategies for Evaluating Saliency Maps in Explainable AI (XAI)
- Seeking Flat Minima over Diverse Surrogates for Improved Adversarial Transferability: A Theoretical Framework and Algorithmic Instantiation
- IAN: Combining Generative Adversarial Networks for Imaginative Face Generation
- Reducing Adversarial Example Transferability Using Gradient Regularization
- Collaboration Analysis Using Deep Learning
- Reducing the Amortization Gap in Variational Autoencoders: A Bayesian Random Function Approach
- Analyzing Learned Convnet Features with Dirichlet Process Gaussian Mixture Models
- Fingertip detection and tracking for recognition of air-writing in videos
- Machine-learning assisted quantum control in random environment
- Identifying eclipsing binary stars with TESS data based on a new hybrid deep learning model
- Deep learning and face recognition: the state of the art
- Deep Visual Waterline Detection within Inland Marine Environment
- Towards Understanding Acceleration Tradeoff between Momentum and Asynchrony in Nonconvex Stochastic Optimization
- Progressive Tandem Learning for Pattern Recognition with Deep Spiking Neural Networks
- TDAsweep: A Novel Dimensionality Reduction Method for Image Classification Tasks
- LayerPipe: Accelerating Deep Neural Network Training by Intra-Layer and Inter-Layer Gradient Pipelining and Multiprocessor Scheduling
- Deep triplet hashing network for case-based medical image retrieval
- CNN-Based Deep Learning Model for Solar Wind Forecasting
- An accelerated correlation filter tracker
- Knowledge Guided Disambiguation for Large-Scale Scene Classification With Multi-Resolution CNNs
- An attention-fused network for semantic segmentation of very-high-resolution remote sensing imagery
- Music Genre Classification Using Masked Conditional Neural Networks
- Neural Rating Regression with Abstractive Tips Generation for Recommendation
- Light-weighted Saliency Detection with Distinctively Lower Memory Cost and Model Size
- Relationship Oriented Affordance Learning through Manipulation Graph Construction
- On the Number of Linear Functions Composing Deep Neural Network: Towards\n a Refined Definition of Neural Networks Complexity
- Latent Cognizance: What Machine Really Learns
- Improving utility of brain tumor confocal laser endomicroscopy: objective value assessment and diagnostic frame detection with convolutional neural networks
- Global convergence of neuron birth-death dynamics
- Adversarial Attacks on Deep Neural Networks for Time Series Classification
- Deep Neural Network Ensembles for Time Series Classification
- Rapid visual categorization is not guided by early salience-based selection
- Learning deformable registration of medical images with anatomical constraints
- Driving Scene Perception Network: Real-Time Joint Detection, Depth Estimation and Semantic Segmentation
- Building medical image classifiers with very limited data using segmentation networks
- Multiview Deep Learning for Predicting Twitter Users' Location
- Complex-valued neural networks for machine learning on non-stationary physical data
- Prospects for Theranostics in Neurosurgical Imaging: Empowering Confocal Laser Endomicroscopy Diagnostics via Deep Learning
- Towards responsible AI for education: Hybrid human-AI to confront the Elephant in the room
- On the impact of selected modern deep-learning techniques to the performance and celerity of classification models in an experimental high-energy physics use case
- Boosting KNNClassifier Performance with Opposition-Based Data Transformation
- Large Hole Image Inpainting With Compress-Decompression Network
- ReMix: Calibrated Resampling for Class Imbalance in Deep learning
- Improving Landslide Detection on SAR Data Through Deep Learning
- Integration of Adversarial Autoencoders With Residual Dense Convolutional Networks for Estimation of Non‐Gaussian Hydraulic Conductivities
- Learning from Few Samples: A Survey
- Adapting a global plant identification model to detect invasive alien plant species in high-resolution road side images
- SSDH: Semi-Supervised Deep Hashing for Large Scale Image Retrieval
- Analytical Softmax Temperature Setting from Feature Dimensions for Model- and Domain-Robust Classification
- The Open Images Dataset V4
- Enabling Spike-Based Backpropagation for Training Deep Neural Network Architectures
- CANet: An Unsupervised Intrusion Detection System for High Dimensional CAN Bus Data
- Deep-learning-based image segmentation integrated with optical microscopy for automatically searching for two-dimensional materials
- Deep learning with missing data
- Autoencoding sensory substitution
- Controlling information capacity of binary neural network
- VeLU: Variance-enhanced Learning Unit for Deep Neural Networks
- Symmetry reduction for deep reinforcement learning active control of chaotic spatiotemporal dynamics
- Locally Supervised Deep Hybrid Model for Scene Recognition
- Language Models for Materials Discovery and Sustainability: Progress, Challenges, and Opportunities
- ECViT: Efficient Convolutional Vision Transformer with Local-Attention and Multi-scale Stages
- Hybrid Knowledge Transfer through Attention and Logit Distillation for On-Device Vision Systems in Agricultural IoT
- Network Quantization with Element-wise Gradient Scaling
- Multiclass wound image classification using an ensemble deep CNN-based classifier
- Can We Ignore Labels In Out of Distribution Detection?
- Efficient Split Federated Learning for Large Language Models over Communication Networks
- Hardware-friendly Neural Network Architecture for Neuromorphic Computing
- Clustering and Classification Networks
- Bangla Handwritten Digit Recognition and Generation
- On Sampling Strategies for Neural Network-based Collaborative Filtering
- Noise-Level Estimation from Single Color Image Using Correlations Between Textures in RGB Channels
- Multiperson Continuous Tracking and Identification From mm-Wave Micro-Doppler Signatures
- Acoustic Anomaly Detection for Machine Sounds based on Image Transfer Learning
- Accelerating Deep Learning with Shrinkage and Recall
- Deep neural network for solving differential equations motivated by Legendre-Galerkin approximation
- Text Detection and Recognition in the Wild: A Review
- Data-driven Regularization via Racecar Training for Generalizing Neural Networks
- StackMix: A complementary Mix algorithm
- Deep Learning-Based Vehicle Behavior Prediction for Autonomous Driving Applications: A Review
- Ensemble Kalman inversion: a derivative-free technique for machine learning tasks
- FingerNet: Pushing The Limits of Fingerprint Recognition Using Convolutional Neural Network
- Watch It Twice: Video Captioning with a Refocused Video Encoder
- Context-based object viewpoint estimation: A 2D relational approach
- Quasi-hyperbolic momentum and Adam for deep learning
- Human Pose and Path Estimation from Aerial Video Using Dynamic Classifier Selection
- Free annotated data for deep learning in microscopy? A hitchhiker’s guide
- Neural ODE to model and prognose thermoacoustic instability
- Convolutional Sparse Coding Fast Approximation With Application to Seismic Reflectivity Estimation
- SimplifyMyText: An LLM-Based System for Inclusive Plain Language Text Simplification
- Occlusion Edge Detection in RGB-D Frames using Deep Convolutional Networks
- Leveraging Generative AI Models to Explore Human Identity
- Internal noise in hardware deep and recurrent neural networks helps with learning
- Rapidly Adapting Moment Estimation
- Networks with pixels embedding: a method to improve noise resistance in images classification
- AutoSlim: Towards One-Shot Architecture Search for Channel Numbers
- A Study on Action Detection in the Wild
- AG-CUResNeSt: A Novel Method for Colon Polyp Segmentation
- Fashionista: A Fashion-aware Graphical System for Exploring Visually Similar Items
- EDropout: Energy-Based Dropout and Pruning of Deep Neural Networks
- Spatially-Adaptive Filter Units for Compact and Efficient Deep Neural Networks
- 3D Human Pose Machines with Self-supervised Learning
- Masked Conditional Neural Networks for Audio Classification
- Shoulder Implant X-Ray Manufacturer Classification: Exploring with Vision Transformer
- Scattering Networks for Hybrid Representation Learning
- Ellipse R-CNN: Learning to Infer Elliptical Object From Clustering and Occlusion
- A survey on incorporating domain knowledge into deep learning for medical image analysis
- Data Augmentation via Mixed Class Interpolation using Cycle-Consistent Generative Adversarial Networks Applied to Cross-Domain Imagery
- ActionFlowNet: Learning Motion Representation for Action Recognition
- Unsupervised Deep Representation Learning and Few-Shot Classification of PolSAR Images
- Robust unsupervised domain adaptation for neural networks via moment alignment
- Joint Facade Registration and Segmentation for Urban Localization
- On-device Filtering of Social Media Images for Efficient Storage
- HPC AI500: The Methodology, Tools, Roofline Performance Models, and Metrics for Benchmarking HPC AI Systems
- A Survey on Machine Learning Techniques for Source Code Analysis
- Urban Anomaly Analytics: Description, Detection, and Prediction
- Temporally Resolution Decrement: Utilizing the Shape Consistency for Higher Computational Efficiency
- Deep ensemble network with explicit complementary model for accuracy-balanced classification
- MAAM: A Lightweight Multi-Agent Aggregation Module for Efficient Image Classification Based on the MindSpore Framework
- OBIFormer: A Fast Attentive Denoising Framework for Oracle Bone Inscriptions
- Simulating Nonlinearity in Quantum Neural Networks While Mitigating Barren Plateaus
- Denoising Prior Driven Deep Neural Network for Image Restoration
- GraphAIR: Graph representation learning with neighborhood aggregation and interaction
- Single image portrait relighting
- Automatic model based dataset generation for fast and accurate crop and weeds detection
- Deep Learning-Based Video Coding
- One-Shot Recognition of Manufacturing Defects in Steel Surfaces
- DASNet: Dynamic Activation Sparsity for Neural Network Efficiency Improvement
- Understanding and Improving Virtual Adversarial Training
- Representing text as abstract images enables image classifiers to also simultaneously classify text
- Analyzing the Dependency of ConvNets on Spatial Information
- Artificial intelligence and deep learning algorithms for epigenetic sequence analysis: A review for epigeneticists and AI experts
- Res2Net: A New Multi-Scale Backbone Architecture
- Detecting The Objects on The Road Using Modular Lightweight Network
- Data-driven rogue waves and parameter discovery in the defocusing nonlinear Schrödinger equation with a potential using the PINN deep learning
- Domain-specific cues improve robustness of deep learning-based segmentation of CT volumes
- BoxCars: Improving Fine-Grained Recognition of Vehicles Using 3-D Bounding Boxes in Traffic Surveillance
- Deep learning techniques for in-crop weed recognition in large-scale grain production systems: a review
- CG-Net: Conditional GIS-aware Network for Individual Building Segmentation in VHR SAR Images
- Learning Without Feedback: Fixed Random Learning Signals Allow for Feedforward Training of Deep Neural Networks
- Image Super-Resolution with Cross-Scale Non-Local Attention and Exhaustive Self-Exemplars Mining
- Hidden unit specialization in layered neural networks: ReLU vs. sigmoidal activation
- Parameter estimation for the cosmic microwave background with Bayesian neural networks
- Magnification Generalization For Histopathology Image Embedding
- An Empirical Evaluation on Robustness and Uncertainty of Regularization Methods
- Pacemaker: Intermediate Teacher Knowledge Distillation For On-The-Fly Convolutional Neural Network
- High-Throughput CNN Inference on Embedded ARM Big.LITTLE Multicore Processors
- Epileptic Seizures Detection Using Deep Learning Techniques: A Review
- Discriminative Feature Learning With Foreground Attention for Person Re-Identification
- Fog Computing on Constrained Devices: Paving the Way for the Future IoT
- Deep learning microscopy
- Enforcing Deterministic Constraints on Generative Adversarial Networks for Emulating Physical Systems
- A Modular Analysis of Provable Acceleration via Polyak's Momentum: Training a Wide ReLU Network and a Deep Linear Network
- An Overview of Computational Approaches for Interpretation Analysis
- Seeing permeability from images: fast prediction with convolutional neural networks
- t-soft update of target network for deep reinforcement learning
- Deep divergence-based approach to clustering
- The Emergence of Canalization and Evolvability in an Open-Ended, Interactive Evolutionary System
- Figuring Out Art History
- Cluster Pruning: An Efficient Filter Pruning Method for Edge AI Vision Applications
- An artificial intelligence atomic force microscope enabled by machine learning
- PaintBot: A Reinforcement Learning Approach for Natural Media Painting
- Deep learning in remote sensing applications: A meta-analysis and review
- Review on Convolutional Neural Networks (CNN) in vegetation remote sensing
- Rational Neural Networks for Approximating Jump Discontinuities of Graph Convolution Operator
- Deep Learning for Sensor-based Human Activity Recognition: Overview, Challenges and Opportunities
- Learning Spectral-Spatial-Temporal Features via a Recurrent Convolutional Neural Network for Change Detection in Multispectral Imagery
- The Application of Artificial Intelligence to Acoustic Data in Otolaryngology
- Application of Multi-channel 3D-cube Successive Convolution Network for Convective Storm Nowcasting
- Learning Material-Aware Local Descriptors for 3D Shapes
- Toward Accurate Platform-Aware Performance Modeling for Deep Neural Networks
- Local Aggregation for Unsupervised Learning of Visual Embeddings
- Decorrelated Adversarial Learning for Age-Invariant Face Recognition
- Automated Search for Configurations of Deep Neural Network Architectures
- Dompteur: Taming Audio Adversarial Examples
- Deep Learning-based Type Identification of Volumetric MRI Sequences
- Complex sequential understanding through the awareness of spatial and temporal concepts
- Photometric search for exomoons by using convolutional neural networks
- Improvements to Context Based Self-Supervised Learning
- Comparing heterogeneous entities using artificial neural networks of trainable weighted structural components and machine-learned activation functions
- Jamming transition as a paradigm to understand the loss landscape of deep neural networks
- Real-time Deep Video Deinterlacing
- Neural Architecture Transfer
- Deep convolutions for in-depth automated rock typing
- Improving Semantic Analysis on Point Clouds via Auxiliary Supervision of Local Geometric Priors
- AxonDeepSeg: automatic axon and myelin segmentation from microscopy data using convolutional neural networks
- Evolution of Image Segmentation using Deep Convolutional Neural Network: A Survey
- Deep Learning for Visual Tracking: A Comprehensive Survey
- Deep model with Siamese network for viable and necrotic tumor regions assessment in osteosarcoma
- Visual and Semantic Knowledge Transfer for Large Scale Semi-Supervised Object Detection
- Deep Learning for Scene Classification: A Survey
- NNWarp: Neural Network-based Nonlinear Deformation
- The Emerging Trends of Multi-Label Learning
- What Looks Good with my Sofa: Ensemble Multimodal Search for Interior Design
- Understanding More about Human and Machine Attention in Deep Neural Networks
- Automatic Prediction of Building Age from Photographs
- Semantic Relation Preserving Knowledge Distillation for Image-to-Image Translation
- SlowFast Networks for Video Recognition
- Tom: Leveraging trend of the observed gradients for faster convergence
- Adversarial Color Enhancement: Generating Unrestricted Adversarial Images by Optimizing a Color Filter
- Characterizing Adversarial Examples Based on Spatial Consistency Information for Semantic Segmentation
- Rice Classification Using Spatio-Spectral Deep Convolutional Neural Network
- Planning Robot Motion using Deep Visual Prediction
- New bag of deep visual words based features to classify chest x-ray images for COVID-19 diagnosis
- Spatial–Temporal Recurrent Neural Network for Emotion Recognition
- Dual Contradistinctive Generative Autoencoder
- Recognition of ischaemia and infection in diabetic foot ulcers: Dataset and techniques
- Deep Learning Architect: Classification for Architectural Design through the Eye of Artificial Intelligence
- Environmental Sound Recognition Using Masked Conditional Neural Networks
- Cardiac MRI Semantic Segmentation for Ventricles and Myocardium using Deep Learning
- Putting the Segment Anything Model to the Test with 3D Knee MRI - A Comparison with State-of-the-Art Performance
- EgoFace: Egocentric Face Performance Capture and Videorealistic\n Reenactment
- A multi-path 2.5 dimensional convolutional neural network system for segmenting stroke lesions in brain MRI images
- ArtistAuditor: Auditing Artist Style Pirate in Text-to-Image Generation Models
- On Procedural Adversarial Noise Attack And Defense
- Geometric deep learning for computational mechanics Part I: anisotropic hyperelasticity
- Survey on Vision-Based Path Prediction
- A Comprehensive Survey of Challenges and Opportunities of Few-Shot Learning Across Multiple Domains
- Deep learning at the shallow end: Malware classification for non-domain experts
- AdaptoVision: A Multi-Resolution Image Recognition Model for Robust and Scalable Classification
- What does fault tolerant Deep Learning need from MPI?
- Quantum Computing Supported Adversarial Attack-Resilient Autonomous Vehicle Perception Module for Traffic Sign Classification
- Classification-Based Analysis of Price Pattern Differences Between Cryptocurrencies and Stocks
- Locally Scale-Invariant Convolutional Neural Networks
- ADA-Net: Attention-Guided Domain Adaptation Network with Contrastive Learning for Standing Dead Tree Segmentation Using Aerial Imagery
- Human counting versus artificial intelligence for assessing medullation in mohair fibres
- TacoDepth: Towards Efficient Radar-Camera Depth Estimation with One-stage Fusion
- Novel-view X-ray Projection Synthesis through Geometry-Integrated Deep Learning
- RDI: An adversarial robustness evaluation metric for deep neural networks based on model statistical features
- Deep Anatomical Federated Network (Dafne): An Open Client-Server Framework for Continuous, Collaborative Improvement of Deep Learning–based Medical Image Segmentation
- Generating Adversarial Examples with an Optimized Quality
- Deep Learning-Based Prediction of Key Performance Indicators for Electrical Machines
- Threshold-Based Early Stopping of Accumulations in Neural Networks with Binary Activation
- Mapping at First Sense: A Lightweight Neural Network-Based Indoor Structures Prediction Method for Robot Autonomous Exploration
- MirBot: A collaborative object recognition system for smartphones using convolutional neural networks
- Topological shape transform for thymus structures
- Classifying and segmenting microscopy images with deep multiple instance learning
- The Pontryagin Maximum Principle for Training Convolutional Neural Networks
- AtlasD: Automatic Local Symmetry Discovery
- SoK: Can Fully Homomorphic Encryption Support General AI Computation? A Functional and Cost Analysis
- ConvShareViT: Enhancing Vision Transformers with Convolutional Attention Mechanisms for Free-Space Optical Accelerators
- The road to commercial success for neuromorphic technologies
- Action-Attending Graphic Neural Network
- An Image is Worth K Topics: A Visual Structural Topic Model with Pretrained Image Embeddings
- Learning with Spike Synchrony in Spiking Neural Networks
- GFT: Gradient Focal Transformer
- EBAD-Gaussian: Event-driven Bundle Adjusted Deblur Gaussian Splatting
- LEMUR Neural Network Dataset: Towards Seamless AutoML
- Continual learning for rotating machinery fault diagnosis with cross-domain environmental and operational variations
- BLAST: Bayesian online change-point detection with structured image data
- FractalForensics: Proactive Deepfake Detection and Localization via Fractal Watermarks
- A review on deep learning for vision-based hand detection, hand segmentation and hand gesture recognition in human–robot interaction
- Bregman Linearized Augmented Lagrangian Method for Nonconvex Constrained Stochastic Zeroth-order Optimization
- B-DRRN: A Block Information Constrained Deep Recursive Residual Network for Video Compression Artifacts Reduction
- Beyond Glucose-Only Assessment: Advancing Nocturnal Hypoglycemia Prediction in Children with Type 1 Diabetes
- Spiking Neural Network for Intra-cortical Brain Signal Decoding
- Quantum Large Language Model Fine-Tuning
- Personalizing Federated Learning for Hierarchical Edge Networks with Non-IID Data
- Light-YOLOv8-Flame: A Lightweight High-Performance Flame Detection Algorithm
- Comparative Analysis of Different Methods for Classifying Polychromatic Sketches
- Boosting-inspired online learning with transfer for railway maintenance
- Statistically guided deep learning
- Domain Generalization via Semi-supervised Meta Learning
- Smoother Network Tuning and Interpolation for Continuous-level Image Processing
- Joint Vertebrae Identification and Localization in Spinal CT Images by Combining Short- and Long-Range Contextual Information
- GenEAva: Generating Cartoon Avatars with Fine-Grained Facial Expressions from Realistic Diffusion-based Faces
- What Makes Multi-modal Learning Better than Single (Provably)
- Multistream Vision Transformer Fusion with Stain-Physics Priors for Reliable Colorectal Histopathology Classification
- Building damage annotation on post-hurricane satellite imagery based on convolutional neural networks
- A Scale Invariant Flatness Measure for Deep Network Minima
- Digital Neuron: A Hardware Inference Accelerator for Convolutional Deep Neural Networks
- Workload-aware Automatic Parallelization for Multi-GPU DNN Training
- Two-stage framework for optic disc localization and glaucoma classification in retinal fundus images using deep learning
- Attributes-aware Visual Emotion Representation Learning
- Robust Classification with Noisy Labels Based on Posterior Maximization
- Neural Signal Compression using RAMAN tinyML Accelerator for BCI Applications
- Compound and Parallel Modes of Tropical Convolutional Neural Networks
- Beyond Moore's Law: Harnessing the Redshift of Generative AI with Effective Hardware-Software Co-Design
- Adam revisited: a weighted past gradients perspective
- The Function Representation of Artificial Neural Network
- A Meaningful Perturbation Metric for Evaluating Explainability Methods
- LIAF-Net: Leaky Integrate and Analog Fire Network for Lightweight and Efficient Spatiotemporal Information Processing
- Examining Joint Demosaicing and Denoising for Single-, Quad-, and Nona-Bayer Patterns
- GIGA: Generalizable Sparse Image-driven Gaussian Humans
- Optuna vs Code Llama: Are LLMs a New Paradigm for Hyperparameter Tuning?
- Defending Deep Neural Networks against Backdoor Attacks via Module Switching
- Identifying implementation bugs in machine learning based image classifiers using metamorphic testing
- Unsupervised learning from videos using temporal coherency deep networks
- Joint Learning of Neural Transfer and Architecture Adaptation for Image Recognition
- Generative Adversarial Networks with Limited Data: A Survey and Benchmarking
- A Nature-Inspired Colony of Artificial Intelligence System with Fast, Detailed, and Organized Learner Agents for Enhancing Diversity and Quality
- Don't Lag, RAG: Training-Free Adversarial Detection Using RAG
- Strawberry Detection using Mixed Training on Simulated and Real Data
- Distinct contributions of functional and deep neural network features to representational similarity of scenes in human brain and behavior. [europepmc]
- Computer-aided detection in chest radiography based on artificial intelligence: a survey. [europepmc]
- Skin Cancer Classification Using Convolutional Neural Networks: Systematic Review. [europepmc]
- A U-Net Deep Learning Framework for High Performance Vessel Segmentation in Patients With Cerebrovascular Disease. [europepmc]
- Artificial intelligence and machine learning in clinical development: a translational perspective. [europepmc]
- Key Topics in Molecular Docking for Drug Design. [europepmc]
- Artificial Intelligence in Lung Cancer Pathology Image Analysis. [europepmc]
- Evaluation of Combined Artificial Intelligence and Radiologist Assessment to Interpret Screening Mammograms. [europepmc]
- 3D Deep Learning on Medical Images: A Review. [europepmc]
- Deep learning encodes robust discriminative neuroimaging representations to outperform standard machine learning. [europepmc]
- DeepTCR is a deep learning framework for revealing sequence concepts within T-cell repertoires. [europepmc]
- Review of deep learning: concepts, CNN architectures, challenges, applications, future directions. [europepmc]
- The language of proteins: NLP, machine learning & protein sequences. [europepmc]
- Epileptic Seizures Detection Using Deep Learning Techniques: A Review. [europepmc]
- GNINA 1.0: molecular docking with deep learning. [europepmc]
- Combining Machine Learning and Computational Chemistry for Predictive Insights Into Chemical Systems. [europepmc]
- TransMed: Transformers Advance Multi-Modal Medical Image Classification. [europepmc]
- A review on deep learning in medical image analysis. [europepmc]
- Recent Advances in Electrochemical Biosensors: Applications, Challenges, and Future Scope. [europepmc]
- The Role of Artificial Intelligence in Early Cancer Diagnosis. [europepmc]
- Radiology artificial intelligence: a systematic review and evaluation of methods (RAISE). [europepmc]
- A fully automatic AI system for tooth and alveolar bone segmentation from cone-beam CT images. [europepmc]
- Explainable medical imaging AI needs human-centered design: guidelines and evidence from a systematic review. [europepmc]
- Redefining Radiology: A Review of Artificial Intelligence Integration in Medical Imaging. [europepmc]
- How Artificial Intelligence Is Shaping Medical Imaging Technology: A Survey of Innovations and Applications. [europepmc]
Related