Exploiting Generative AI to Scale up Intelligent Tutoring Systems
2023/01/01 by Jakubův, Jan, Chvalovský, Karel, Goertzel, Zarathustra +6 · 304 citations
Computer Science · #Multimodal Machine Learning Applications #Natural Language Processing Techniques #Topic Modeling
paper · pdf · doi:10.4230/lipics.itp.2023.19
openalex publication_date 2023/01/01 · openalex created_date 2023/07/26 · openalex updated_date 2026/07/30
Abstract
As a present to Mizar on its 50th anniversary, we develop an AI/TP system that automatically proves about 60% of the Mizar theorems in the hammer setting. We also automatically prove 75% of the Mizar theorems when the automated provers are helped by using only the premises used in the human-written Mizar proofs. We describe the methods and large-scale experiments leading to these results. This includes in particular the E and Vampire provers, their ENIGMA and Deepire learning modifications, a number of learning-based premise selection methods, and the incremental loop that interleaves growing a corpus of millions of ATP proofs with training increasingly strong AI/TP systems on them. We also present a selection of Mizar problems that were proved automatically.
Cited by
- A ConvNet for the 2020s
- MetaFormer is Actually What You Need for Vision
- Next word prediction for Urdu language using deep learning models
- Pyramid Vision Transformer: A Versatile Backbone for Dense Prediction without Convolutions
- CrossViT: Cross-Attention Multi-Scale Vision Transformer for Image Classification
- Immersive exposure to simulated visual hallucinations modulates high-level human cognition
- Gender biases and hate speech: Promoters and targets in the Argentinean political context
- Learning skillful medium-range global weather forecasting
- A Survey on Deep Learning for Named Entity Recognition
- A Universal Subhypergraph-Assisted Embedding Framework for Both Homogeneous and Heterogeneous Networks
- S-MGHSTN: Towards An Effective Streaming Traffic Accident Risk Prediction Framework
- High-Resolution Image Synthesis with Latent Diffusion Models
- Skillful Precipitation Nowcasting using Deep Generative Models of Radar
- Swin Transformer: Hierarchical Vision Transformer using Shifted Windows
- Learning defects from aircraft NDT data
- Acoustic source localization by deep-learning attention-based modulation of microphone array data
- How far can you go? Extrapolating values of catalytic activity from known protein landscapes in natural and directed evolution
- ProGen2: Exploring the boundaries of protein language models
- Large Language Models and Generative AI, Oh My!
- Recent Advances in Named Entity Recognition: A Comprehensive Survey and Comparative Study
- Deep Graph Memory Networks for Forgetting-Robust Knowledge Tracing
- Effective and Efficient Multi-View Imputation With Optimal Transport
- Exclusive: the most-cited papers of the twenty-first century
- Vision Transformers for X-ray Diffraction Patterns Analysis
- Large Language Models and the Future of Organization Theory
- The TESCREAL bundle: Eugenics and the promise of utopia through artificial general intelligence
- How to Use Generative AI in Educational Research
- Remembering without (representational) memory: a neuro-computational study on regaining categoricity and compositionality from minimal traces
- The generative era of medical AI
- Methylomes Reveal Recent Evolutionary Changes in Populations of Two Plant Species
- Cyberdelics: Virtual reality hallucinations modulate cognitive-affective processes
- Benchmarking DNA large language models on quadruplexes
- Pixels and Predictions: Potential of GPT-4V in Meteorological Imagery Analysis and Forecast Communication
- Revealing Rubric Relations: Investigating the Interdependence of a Research-Informed and a Machine Learning-Based Rubric in Assessing Student Reasoning in Chemistry
- Multimodal mixing convolutional neural network and transformer for Alzheimer’s disease recognition
- Heterogeneous multivariate time series imputation by transformer model with missing position encoding
- Short-term photovoltaic power forecasting with feature extraction and attention mechanisms
- QPIC: Query-Based Pairwise Human-Object Interaction Detection with Image-Wide Contextual Information
- U-net with ResNet-34 backbone for dual-polarized C-band baltic sea-ice SAR segmentation
- CSwin-PNet: A CNN-Swin Transformer combined pyramid network for breast lesion segmentation in ultrasound images
- Hidformer: Hierarchical dual-tower transformer using multi-scale mergence for long-term time series forecasting
- A deep learning model for multi-modal spatio-temporal irradiance forecast
- Latent diffusion model for conditional reservoir facies generation
- Can we Trust Chatbots for now? Accuracy, reproducibility, traceability; a Case Study on Leonardo da Vinci's Contribution to Astronomy
- Leveraging large language models to predict antibody biological activity against influenza A hemagglutinin
- Sequential Recommendation with Graph Neural Networks
- A Multimodal Deep Learning Approach for Soil Moisture Downscaling Using Remote Sensing and Weather Data
- SwinCrack: Pavement crack detection using convolutional swin-transformer network
- A Survey of Fake News: Fundamental Theories, Detection Methods, and Opportunities
- nnFormer: Volumetric Medical Image Segmentation via a 3D Transformer
- TOPIQ: A Top-Down Approach From Semantics to Distortions for Image Quality Assessment
- Quotation Recommendation and Interpretation Based on Transformation from Queries to Quotations
- Rethinking Semantic Segmentation from a Sequence-to-Sequence Perspective with Transformers
- RGAnomaly: Data reconstruction-based generative adversarial networks for multivariate time series anomaly detection in the Internet of Things
- What Large Language Models Know
- Out of Context: How important is Local Context in Neural Program Repair?
- End-to-end Temporal Action Detection with Transformer
- Fuzzy-ViT: A Deep Neuro-Fuzzy System for Cross-Domain Transfer Learning From Large-Scale General Data to Medical Image
- An Energy-Efficient GeMM-Based Convolution Accelerator With On-the-Fly im2col
- Diffusion Models in Vision: A Survey
- SimCSE: Simple Contrastive Learning of Sentence Embeddings
- Complex business ecosystem intelligence using AI-powered visual analytics
- Latent-KalmanNet: Learned Kalman Filtering for Tracking From High-Dimensional Signals
- Recognition of European mammals and birds in camera trap images using deep neural networks
- A Review of Uncertainty Quantification in Deep Learning: Techniques, Applications and Challenges
- Learning to Prompt for Vision-Language Models
- A Mathematical Investigation of Hallucination and Creativity in GPT Models
- Good for Misconceived Reasons: An Empirical Revisiting on the Need for Visual Context in Multimodal Machine Translation
- E-DSSR: Efficient Dynamic Surgical Scene Reconstruction with Transformer-based Stereoscopic Depth Perception
- Out-of-distribution generalization via composition: A lens through induction heads in Transformers
- CSWin Transformer: A General Vision Transformer Backbone with Cross-Shaped Windows
- Large language models and their applications in bioinformatics
- Image Quality Assessment using Contrastive Learning
- FocalTransNet: A Hybrid Focal-Enhanced Transformer Network for Medical Image Segmentation
- Mapping the unseen in practice: comparing latent Dirichlet allocation and BERTopic for navigating topic spaces
- learnMSA2: deep protein multiple alignments with large language and hidden Markov models
- On Masked Pre-training and the Marginal Likelihood
- Unifying Large Language Models and Knowledge Graphs: A Roadmap
- Collaborative Forensic Autopsy Documentation and Supervised Report Generation Using a Hybrid Mixed-Reality Environment and Generative AI
- Effectiveness of retrieval augmented generation-based large language models for generating construction safety information
- Audio Mamba: Bidirectional State Space Model for Audio Representation Learning
- Swin Transformer V2: Scaling Up Capacity and Resolution
- Cross-disciplinary perspectives on the potential for artificial intelligence across chemistry
- Adapting a global plant identification model to detect invasive alien plant species in high-resolution road side images
- SqueezeCall: nanopore basecalling using a Squeezeformer network
- Are protein language models the new universal key?
- Latent Space Probing for Adult Content Detection in Video Generative Models
- Video Swin Transformer
- CerviFormer: A pap smear‐based cervical cancer classification method using cross‐attention and latent transformer
- ChatGPT for good? On opportunities and challenges of large language models for education
- A Survey on Evaluation of Large Language Models
- BinaryBERT: Pushing the Limit of BERT Quantization
- A Comprehensive Survey of Dynamic Graph Neural Networks: Models, Frameworks, Benchmarks, Experiments and Challenges
- SingSong: Generating musical accompaniments from singing
- Structure-Preserving Transformers for Sequences of SPD Matrices
- Variational Open-Domain Question Answering
- Diffusion tensor estimation with transformer neural networks
- A Mechanistic Interpretation of Arithmetic Reasoning in Language Models using Causal Mediation Analysis
- Data Science in the Big Data Era: Analytics, Intelligence, and Future Challenges
- Multistep traffic forecasting by dynamic graph convolution: Interpretations of real-time spatial correlations
- Human autonomy with AI in the loop
- DeepCodon: A deep learning codon-optimization model to enhance protein expression
- Generating Meaning: Active Inference and the Scope and Limits of Passive AI
- Clinical-grade AI model for molecular subtyping of endometrial cancer: a multi-center cohort study in China
- The Open Science of Deep Learning: Three Case Studies
- Dataset Balancing Can Hurt Model Performance
- Conformer: Convolution-augmented Transformer for Speech Recognition
- Uformer: A General U-Shaped Transformer for Image Restoration
- Time Series Forecasting With Deep Learning: A Survey
- Case reports unlocked: Harnessing large language models to advance research on child maltreatment
- Digital authenticity: Towards a research agenda for the AI-driven fifth phase of digitalization in business-to-business marketing
- Human heuristics for AI-generated language are flawed
- What Are We Automating? On the Need for Vision and Expertise When Deploying AI Systems
- The problem of alignment
- AI Methods for Antimicrobial Peptides: Progress and Challenges
- SparsePoser: Real-time Full-body Motion Reconstruction from Sparse Data
- Generative AI Solutions to Empower Financial Firms
- Privacy and Security Concerns in Generative AI: A Comprehensive Survey
- Using large language models to extract plant functional traits from unstructured text
- From text to traits: exploring the role of large language models in plant breeding
- A systematic evaluation of Dutch large language models’ surprisal estimates in sentence, paragraph and book reading
- Artificial intelligence in horticulture
- MixingDTA: improved drug–target affinity prediction by extending mixup with guilt-by-association
- Machine learning-driven breakthroughs in water electrolysis and supercapacitors
- An Improved Framework for Scaling Party Positions from Texts with Transformer
- Modelling and design of transcriptional enhancers
- Foundation models in bioinformatics
- Time series predictions in unmonitored sites: a survey of machine learning techniques in water resources
- Protocol paper: From Chaos to Order. Augmenting Manual Article Screening with Sentence Transformers in Management Systematic Reviews
- Beyond the Turing Test: Exploring the implications of generative AI for category construction
- Structure in Deep Reinforcement Learning: A Survey and Open Problems
- GPT (Generative Pre-Trained Transformer)— A Comprehensive Review on Enabling Technologies, Potential Applications, Emerging Challenges, and Future Directions
- Archetypal crop trait dynamics for enhanced retrieval of biophysical parameters from Sentinel-2 MSI
- Theoretical Limitations of Self-Attention in Neural Sequence Models
- Bioeconomy firms and where to find them
- Grounded language acquisition through the eyes and ears of a single child
- Low-cost, autonomous microscopy using deep learning and robotics: A crystal morphology case study
- Generative power of a protein language model trained on multiple sequence alignments
- Advancing Automated Content Analysis for a New Era of Media Effects Research: The Key Role of Transfer Learning
- Facing & mitigating common challenges when working with real-world data: The Data Learning Paradigm
- Can ChatGPT pass Glycobiology?
- AI-textuality: Expanding intertextuality to theorize human-AI interaction with generative artificial intelligence
- Using Emotion Embeddings to Transfer Knowledge Between Emotions, Languages, and Annotation Formats
- How fine can fine-tuning be? Learning efficient language models
- Governance of Generative AI
- How to apply zero‐shot learning to text data in substance use research: An overview and tutorial with media data
- Transformer-based audio-visual multimodal fusion for fine-grained recognition of individual sow nursing behaviour
- EDVR: Video Restoration with Enhanced Deformable Convolutional Networks
- MTKGR: multi-task knowledge graph reasoning for food and ingredient recognition
- CodeBERT: A Pre-Trained Model for Programming and Natural Languages
- FFM ‐ ViT : an efficient fish species classification method based on deep features and transformers
- Key-value memory in the brain
- Fusing theory-guided machine learning and bio-sensing: considering time in how children learn science from dynamic multimedia
- Large language models can segment narrative events similarly to humans
- GCT: A Granger-Causal Transformer for Multivariate Traffic Analysis in Smart Villages
- When and how to disclose AI use in academic publishing: AMEE Guide No.192
- A systematic literature review of artificial intelligence (AI) in coaching: insights for future research and product development
- UNetFormer: A UNet-like Transformer for Efficient Semantic Segmentation of Remote Sensing Urban Scene Imagery
- Comparison of traditional machine learning and neural network approaches for automated scoring of second language English essays
- Is Genre Enough? A Theory of Genre Signaling as Generative AI Rhetoric
- Leveraging learned representations and multitask learning for lysine methylation site discovery
- Small, open-source text-embedding models as substitutes to OpenAI models for gene analysis
- Computational nanobody design through deep generative modeling and epitope landscape profiling
- GPS: Harnessing data fusion strategies to improve the accuracy of machine learning-based genomic and phenotypic selection
- Flashzoi: An enhanced Borzoi for accelerated genomic analysis
- Symbol ungrounding: what the successes (and failures) of large language models reveal about human cognition
- Autonomy 2.0: The Quest for Economies of Scale
- A large language model-based agent for wayfinding: simulation of spatial perception and memory
- A scoping review of ChatGPT research in accounting and finance
- Using artificial intelligence to document the hidden RNA virosphere
- Language in the Godless Age of AI
- Evaluation of predictions of disordered binding regions in the CAID2 experiment
- SASD: Self-Attention for Small Datasets—A case study in smart villages
- TransHLA: a Hybrid Transformer model for HLA-presented epitope detection
- Transformers and genome language models
- Structural network measures reveal the emergence of heavy-tailed degree distributions in lottery ticket multilayer perceptrons
- Enhancing predictions of protein stability changes induced by single mutations using MSA-based language models
- Security and Privacy Challenges of Large Language Models: A Survey
- Deep learning and generative artificial intelligence in aging research and healthy longevity medicine
- Conducting Qualitative Interviews with AI
- Leveraging mRNA technology for antigen based immuno-oncology therapies
- SCHA-VAE: Hierarchical Context Aggregation for Few-Shot Generation
- HybridDBRpred: improved sequence-based prediction of DNA-binding amino acids using annotations from structured complexes and disordered proteins
- Language model-based B cell receptor sequence embeddings can effectively encode receptor specificity
- Enhanced prediction of vegetation responses to extreme drought using deep learning and Earth observation data
- Adversarial machine learning :
- Optimizing protein tokenization: reduced amino acid alphabets for efficient and accurate protein language models
- Assessing political bias in AI systems: a framework for disentangling viewpoint preferences from epistemic integrity
- From unannotated to annotated
- CamemBERT: a Tasty French Language Model
- OA-CNNs: Omni-Adaptive Sparse CNNs for 3D Semantic Segmentation
- Protein engineering in the deep learning era
- Deep Learning for 3D Point Clouds: A Survey
- Replay and Synthetic Speech Detection with Res2net Architecture
- The Programmer’s Assistant: Conversational Interaction with a Large Language Model for Software Development
- Deep Learning the Functional Renormalization Group
- VGG-TSwinformer: Transformer-based deep learning model for early Alzheimer’s disease prediction
- Opportunities and Challenges in Applying AI to Evolutionary Morphology
- A Multilevel Multimodal Fusion Transformer for Remote Sensing Semantic Segmentation
- TTSR: A Transformer-Based Topography Neural Network for Digital Elevation Model Super-Resolution
- Recent advances on federated learning: A systematic survey
- Guided Attention for Interpretable Motion Captioning
- Vision-Language Models for Vision Tasks: A Survey
- Aggregated Mutual Learning between CNN and Transformer for semi-supervised medical image segmentation
- SPMFormer: Simplified Physical Model-based transformer with cross-space loss for underwater image enhancement
- Adaptive representation-aligned modeling for visual tracking
- Swin Transformer-Based Multiscale Attention Model for Landslide Extraction From Large-Scale Area
- MMD-MLP: LiDAR-Guided Hyperspectral Data Classification Using Local–Global Directional-MLP With Multiresolution Multiscale Representation
- Learning Effective Representations for Person-Job Fit by Feature Fusion
- Discriminative Adversarial Search for Abstractive Summarization
- Artificial neural networks for neuroscientists: A primer
- Weakly-supervised Fingerspelling Recognition in British Sign Language\n Videos
- DeltaConv: Anisotropic Operators for Geometric Deep Learning on Point Clouds
- Literaturwissenschaft und Informatik
- An Attention-Based Deep Learning Approach for Sleep Stage Classification With Single-Channel EEG
- Shortcut learning in deep neural networks
- Tokens-to-Token ViT: Training Vision Transformers from Scratch on ImageNet
- Raster‐to‐Graph: Floorplan Recognition via Autoregressive Graph Prediction with an Attention Transformer
- A time-series neural network for pig feeding behavior recognition and dangerous detection from videos
- Deep learning in multiple animal tracking: A survey
- Exploiting BERT for End-to-End Aspect-based Sentiment Analysis
- BNAI, NO-TOKEN, and MIND-UNITY: Pillars of a Systemic Revolution in Artificial Intelligence
- Scalable Coupling of Deep Learning with Logical Reasoning
- A Practical Deep Learning-Based Acoustic Side Channel Attack on Keyboards
- Point Transformer
- Goal-conditioned dual-action imitation learning for dexterous dual-arm robot manipulation
- Motion Retargetting based on Dilated Convolutions and Skeleton‐specific Loss Functions
- Medical Image Segmentation Review: The Success of U-Net
- Madness, Cannibalism, and Traditional Fiction between Lu Xun and Mo Yan
- Foundation models for electrocardiogram interpretation: clinical implications
- Arabic Automatic Speech Recognition: Challenges and Progress
- Transformer-based Self-supervised Multimodal Representation Learning for Wearable Emotion Recognition
- Informer: Beyond Efficient Transformer for Long Sequence Time-Series Forecasting
- Development and Validation of a Natural Language Processing Algorithm to Pseudonymize Documents in the Context of a Clinical Data Warehouse
- EVA-02: A visual representation for neon genesis
- Practical and ethical challenges of large language models in education: A systematic scoping review
- Thinking Fast and Slow: Efficient Text-to-Visual Retrieval with Transformers
- Beyond Individual Input for Deep Anomaly Detection on Tabular Data
- Language Conditioned Spatial Relation Reasoning for 3D Object Grounding
- An algorithmic framework for the optimization of deep neural networks architectures and hyperparameters
- One-shot Unsupervised Domain Adaptation with Personalized Diffusion Models
- CALYPSO: LLMs as Dungeon Master's Assistants
- Balsa: Learning a Query Optimizer Without Expert Demonstrations
- Experimental Standards for Deep Learning in Natural Language Processing Research
- Incremental processing of noisy user utterances in the spoken language understanding task
- Artificial Intelligence in Archaeological Geophysics: Testing ChatGPT‐4o on Magnetometer Data From Sapallitepa (Uzbekistan)
- PVT v2: Improved baselines with pyramid vision transformer
- DiffuRec: A Diffusion Model for Sequential Recommendation
- Mixed-Lingual Pre-training for Cross-lingual Summarization
- Multimodal Attention-based Deep Learning for Alzheimer's Disease Diagnosis
- VulTriNet: A software vulnerability detection method based on tri-channel network
- A software vulnerability detection method based on multi-modality with unified processing
- An Attention-based Graph Neural Network for Heterogeneous Structural Learning
- The benefits and dangers of anthropomorphic conversational agents
- An antimicrobial drug recommender system using MALDI-TOF MS and dual-branch neural networks
- Predicting stroke outcome: A case for multimodal deep learning methods with tabular and CT Perfusion data
- Evaluation of BERT and ALBERT Sentence Embedding Performance on Downstream NLP Tasks
- From the Token to the Review: A Hierarchical Multimodal approach to Opinion Mining
- Are Transformers Effective for Time Series Forecasting?
- PICK: Processing Key Information Extraction from Documents using Improved Graph Learning-Convolutional Networks
- Progressive Open-Domain Response Generation with Multiple Controllable Attributes
- Delta Keyword Transformer: Bringing Transformers to the Edge through Dynamically Pruned Multi-Head Self-Attention
- Hybrid semantics-based vulnerability detection incorporating a Temporal Convolutional Network and Self-attention Mechanism
- VAERHNN: Voting-averaged ensemble regression and hybrid neural network to investigate potent leads against colorectal cancer
- Navigating the Complexity of Generative AI Adoption in Software Engineering
- Continuous Entity Reasoning for Multi-Turn Medical Dialog Generation
- Plot and Rework: Modeling Storylines for Visual Storytelling
- A Gentle Introduction to Deep Learning for Graphs
- Intrinsic Dimension Estimation for Robust Detection of AI-Generated Texts
- Introducing AIRSim: An Innovative AI-Driven Feedback Generation Tool for Supporting Student Learning
- Critique of impure reason: Unveiling the reasoning behaviour of medical large language models
- Could generative AI become a ghostwriter for the US president?
- Predictive modeling, pattern recognition, and spatiotemporal representations of plant growth in simulated and controlled environments: A comprehensive review
- Introducing various Semantic Models for Amharic: Experimentation and Evaluation with multiple Tasks and Datasets
- Structured Pruning for Deep Convolutional Neural Networks: A Survey
- A Comprehensive Survey on Source-Free Domain Adaptation
- Scaling Spike-Driven Transformer With Efficient Spike Firing Approximation Training
- Improving Automated Program Repair with Domain Adaptation
- A time series forecasting method for oil production based on Informer optimized by Bayesian optimization and the hyperband algorithm (BOHB)
- Neural Machine Translation with Monolingual Translation Memory
- A novel temporal convolutional network with residual self-attention mechanism for remaining useful life prediction of rolling bearings
- On Inductive Biases for Machine Learning in Data Constrained Settings
- A Survey on Knowledge Graph-Based Recommender Systems
- ChatGPT in healthcare: A taxonomy and systematic review
- Neuron Interaction Based Representation Composition for Neural Machine Translation
- Side-Scan Sonar Underwater Target Detection: Combining the Diffusion Model With an Improved YOLOv7 Model
- HKCoral: Benchmark for Dense Coral Growth Form Segmentation in the Wild
- Self-Training Sampling with Monolingual Data Uncertainty for Neural Machine Translation
- CSLP-AE: A Contrastive Split-Latent Permutation Autoencoder Framework for Zero-Shot Electroencephalography Signal Conversion
- Emerging Properties in Self-Supervised Vision Transformers
- Efficient Diffusion Model for Image Restoration by Residual Shifting
- Transformers: State-of-the-Art Natural Language Processing
- MISSFormer: An Effective Transformer for 2D Medical Image Segmentation
- Gradient scarcity with Bilevel Optimization for Graph Learning
- Controllable Anime Image Editing via Probability of Attribute Tags
- Emotional voice conversion: Theory, databases and ESD
- Text-Independent Speaker Verification with Dual Attention Network
- Voxel Transformer for 3D Object Detection
- Attention-based fusion network for RGB-D semantic segmentation
- Reversible and irreversible bracket-based dynamics for deep graph neural networks
- TPH-YOLOv5: Improved YOLOv5 Based on Transformer Prediction Head for Object Detection on Drone-captured Scenarios
- Augmenting commit classification by using fine-grained source code changes and a pre-trained deep neural language model
- Restormer: Efficient Transformer for High-Resolution Image Restoration
- DiffSBR: A diffusion model for session-based recommendation
- ConDiff: Conditional graph diffusion model for recommendation
- Multi-Scale Transformers with dual attention and adaptive masking for sequential recommendation
- Leveraging Label Correlations in a Multi-label Setting: A Case Study in Emotion
- Run, Don't Walk: Chasing Higher FLOPS for Faster Neural Networks
- Generative artificial intelligence and evaluating strategic decisions
- A multi-task encoder-dual-decoder framework for mixed frequency data prediction
- Interpretable water level forecaster with spatiotemporal causal attention mechanisms
- Asymptotic theory of in-context learning by linear attention
- Generative AI for Software Practitioners
- Sense Vocabulary Compression through the Semantic Knowledge of WordNet for Neural Word Sense Disambiguation
- Reconstructing hadronically decaying tau leptons with a jet foundation model
- Partial Discharge Identification Based on Unsupervised Representation Learning Under Repetitive Impulse Excitation With Ultra-Fast Slew-Rate
- A HO-BiGRU-Transformer based PEMFC degradation prediction method under different current conditions
- Multi-scale hybrid vision transformer and Sinkhorn tokenizer for sewer defect classification
- A Hierarchical Neural Framework for Classification and its Explanation in Large Unstructured Legal Documents
- “Desired behaviors”: alignment and the emergence of a machine learning ethics
- Explainable modeling of single-cell perturbation data using attention and sparse dictionary learning
- Relevance-Promoting Language Model for Short-Text Conversation
- GMAN: A Graph Multi-Attention Network for Traffic Prediction
- A Survey of Human-in-the-loop for Machine Learning
- Deep learning-based waste detection in natural and urban environments
- Mathematical Word Problem Generation from Commonsense Knowledge Graph and Equations
- A Robust Image Semantic Communication System With Multi-Scale Vision Transformer
- Compute Trends Across Three Eras of Machine Learning
- MSFA-Net: Multi-scale feature aggregation and attention-enhanced U-Net for microscopic hyperspectral pathology images segmentation
- Social physics
- Graph Neural Network for Traffic Forecasting: A Survey
- Domain-Specific Language Model Pretraining for Biomedical Natural Language Processing
- Image Representations Learned With Unsupervised Pre-Training Contain Human-like Biases
- How to build the virtual cell with artificial intelligence: Priorities and opportunities
- The effects of over-reliance on AI dialogue systems on students' cognitive abilities: a systematic review
- Transforming Physiology and Healthcare through Foundation Models
- Chain of Risks Evaluation (CORE): A framework for safer large language models in public mental health
- Cross-Media Keyphrase Prediction: A Unified Framework with Multi-Modality Multi-Head Attention and Image Wordings
- Misspelling Correction with Pre-trained Contextual Language Model
- A novel scene coupling semantic mask network for remote sensing image segmentation
- Metaheuristics and Large Language Models Join Forces: Toward an Integrated Optimization Approach
- Regression as Classification: Influence of Task Formulation on Neural Network Features
- Online Learning Meets Machine Translation Evaluation: Finding the Best Systems with the Least Human Effort
- Backpropagation through time and the brain
- Retrieval-based Goal-Oriented Dialogue Generation
- From Multimodal to Unimodal Attention in Transformers using Knowledge Distillation
- Fast Unsupervised Deep Outlier Model Selection with Hypernetworks
- Quotation Recommendation for Multi-party Online Conversations Based on Semantic and Topic Fusion
- StyleTTS 2: Towards Human-Level Text-to-Speech through Style Diffusion and Adversarial Training with Large Speech Language Models
- PIP-Net: Pedestrian Intention Prediction in the Wild
- U-Net and its variants for medical image segmentation: theory and applications
- Developing and assessing second language listening and speaking: Does AI make it better?
- Multivariate image processing in minerals engineering with vision transformers
- Prompt engineering for bibliographic web-scraping
- META4: Semantically-Aligned Generation of Metaphoric Gestures Using Self-Supervised Text and Speech Representation
- Motor imagery EEG decoding based on TS-former for spinal cord injury patients
- Generative AI and the social functions of educational assessment
- The pointer network for reward maximisation in multi-target space mission sequence selection
- Inductive Representation Learning in Temporal Networks via Causal Anonymous Walks
- How can we make robot dance expressive and responsive? A survey of methods and future directions
- AlignTTS: Efficient Feed-Forward Text-to-Speech System without Explicit Alignment
- Learning Multiscale Correlations for Human Motion Prediction
- WavLM: Large-Scale Self-Supervised Pre-Training for Full Stack Speech Processing
- Application of AI in biological age prediction
- Rewrite the Stars
- Global insights and the impact of generative AI-ChatGPT on multidisciplinary: a systematic review and bibliometric analysis
- KoreALBERT: Pretraining a Lite BERT Model for Korean Language Understanding
- A case study of forensic psychiatry experts' reports analysis through large language models
- Accurate Label Refinement From Multiannotator of Remote Sensing Data
- A Unified Framework With Multimodal Fine-Tuning for Remote Sensing Semantic Segmentation
- Mapping Countrywide Historical Tree Cover Using Semantic Segmentation
- Reliable generation of privacy-preserving synthetic electronic health record time series via diffusion models
- Multi-Head Adapter Routing for Cross-Task Generalization
- Physics-Informed machine learning for solar-thermal power systems
- AfriSign: African sign languages machine translation
- Competition between AI foundation models: dynamics and policy recommendations
- Artificial intelligence without restriction surpassing human intelligence with probability one: Theoretical insight into secrets of the brain with AI twins of the brain
- Intelligent summaries: Will Artificial Intelligence mark the finale for biomedical literature reviews?
- Protein–protein and protein–nucleic acid binding site prediction via interpretable hierarchical geometric deep learning
- RNA language models predict mutations that improve RNA function
- Emotion topology: extracting fundamental components of emotions from text using word embeddings
- Drug target prediction through deep learning functional representation of gene signatures
- Zero shot health trajectory prediction using transformer
- Thalamocortical architectures for flexible cognition and efficient learning
- Human Versus Machine Intelligence: Assessing Natural Language Generation Models Through Complex Systems Theory
- From Angels to Artificial Agents? AI as a Mirror for Human (Im)perfections
- Towards a machine-learned Poisson solver for low-temperature plasma simulations in complex geometries
- La nueva realidad de la educación ante los avances de la inteligencia artificial generativa
- A Survey on Deep Neural Network Pruning: Taxonomy, Comparison, Analysis, and Recommendations
- DiffusionAD: Norm-guided One-step Denoising Diffusion for Anomaly Detection
- Short-term Hebbian learning can implement transformer-like attention
- Why Literary Translators should embrace Translation Technology
- MultiCycGT: A Deep Learning-Based Multimodal Model for Predicting the Membrane Permeability of Cyclic Peptides
- Advancing Accessibility through Rigorous Mathematical Models for Cross-Sensory Translation
- Towards a safe and efficient clinical implementation of machine learning in radiation oncology by exploring model interpretability, explainability and data-model dependency
- Joint Intent Detection And Slot Filling Based on Continual Learning Model
- SwinIR: Image Restoration Using Swin Transformer
- Language Modelling with Pixels
- A Hybrid Deterministic Framework for Named Entity Extraction in Broadcast News Video
- Improving First-stage Retrieval of Point-of-interest Search by Pre-training Models
- A Survey on Aspect-Based Sentiment Analysis: Tasks, Methods, and Challenges
- A Survey on Generative Diffusion Models
- Beyond Co-Occurrence: Multi-Modal Session-Based Recommendation
- A pragmatic and intelligent model for sarcasm detection in social media text
- Applications of deep learning in stock market prediction: recent progress
- Using computational modeling to validate the onset of productive determiner–noun combinations in English-learning children
- Achieving generalized three-dimensional flow field prediction for high-speed flight vehicles using an attention-inspired architecture
- A Glitch in the Matrix? Locating and Detecting Language Model Grounding with Fakepedia
- H2CGL: Modeling dynamics of citation network for impact prediction
- Environment Semantics Aided Wireless Communications: A Case Study of mmWave Beam Prediction and Blockage Prediction
- The limits of large language models and the necessity of human cognition in K-12 education
- Transformer versus traditional natural language processing: how much data is enough for automated radiology report classification?
- ForceGen: End-to-end de novo protein generation based on nonlinear mechanical unfolding responses using a protein language diffusion model
- Minimal Neural Network Models for Permutation Invariant Agents
- ESCOXLM-R: Multilingual Taxonomy-driven Pre-training for the Job Market Domain
- Adverse drug events and medication relation extraction in electronic health records with ensemble deep learning methods
- Incorporating User Generated Content for Drug Drug Interaction Extraction Based on Full Attention Mechanism
- Distillation-based fabric anomaly detection
- Reconciling deep learning with symbolic artificial intelligence: representing objects and relations
- TL-med: A Two-stage transfer learning recognition model for medical images of COVID-19
- VQ-HPS
- The Artificial Intelligence Assessment Scale (AIAS): A Framework for Ethical Integration of Generative AI in Educational Assessment
- Revisiting Neural Retrieval on Accelerators
- Micro-Segmentation Anomaly Detection in Zero-Trust Software-Defined Network Fabrics
- Heads-up! Unsupervised Constituency Parsing via Self-Attention Heads
- Adversarial Stylometry in the Wild: Transferable Lexical Substitution Attacks on Author Profiling
- HSNet: Crowd counting via hierarchical scale calibration and spatial attention
- Low resource recognition and linking of biomedical concepts from a large ontology
- A Hybrid Learning Method for System Identification and Optimal Control
- STAN: Spatio-Temporal Attention Network for Next Location Recommendation
- A General Survey on Attention Mechanisms in Deep Learning
- DepMSTAT: Multimodal Spatio-Temporal Attentional Transformer for Depression Detection
- Next-POI Recommendation via Spatial-Temporal Knowledge Graph Contrastive Learning and Trajectory Prompt
- Towards an Online Empathetic Chatbot with Emotion Causes
- Representation Learning for Natural Language Processing
- Inf-VAE
- Attention-based Clinical Note Summarization
- Spinning Language Models: Risks of Propaganda-As-A-Service and Countermeasures
- R2D2: Recursive Transformer based on Differentiable Tree for Interpretable Hierarchical Language Modeling
- A Systematic Review of the Use of Deep Learning in Satellite Imagery for Agriculture
- Ovid: A Machine Learning Approach for Automated Vandalism Detection in OpenStreetMap
- Annotating Columns with Pre-trained Language Models
- RPT: Toward Transferable Model on Heterogeneous Researcher Data via Pre-Training
- Speech Emotion Recognition Via CNN-Transformer and Multidimensional Attention Mechanism
- It’s Morphin’ Time! Combating Linguistic Discrimination with Inflectional Perturbations
- Will We Ever Have Conscious Machines?
- Simultaneous imputation and disease classification in incomplete medical datasets using Multigraph Geometric Matrix Completion (MGMC)
- PhenoTagger: a hybrid method for phenotype concept recognition using human phenotype ontology
- When FastText Pays Attention: Efficient Estimation of Word Representations using Constrained Positional Weighting
- Deep triplet hashing network for case-based medical image retrieval
- A Single-Shot Arbitrarily-Shaped Text Detector based on Context Attended Multi-Task Learning
- Understanding neural code intelligence through program simplification
- Knowledge-Preserving Incremental Social Event Detection via Heterogeneous GNNs
- Is Graph Structure Necessary for Multi-hop Question Answering?
- Harnessing Evolution of Multi-Turn Conversations for Effective Answer Retrieval
- Temporal Sequence Distillation
- Anomaly Detection on Attributed Networks via Contrastive Self-Supervised Learning
- Multi-Modal Sarcasm Detection and Humor Classification in Code-Mixed Conversations
- Fine-tuning ERNIE for chest abnormal imaging signs extraction
- CNN Attention Guidance for Improved Orthopedics Radiographic Fracture Classification
- StackRec
- DASNet: Dual Attentive Fully Convolutional Siamese Networks for Change Detection in High-Resolution Satellite Images
- Automated data processing and feature engineering for deep learning and big data applications: A survey
- Prior Knowledge Driven Label Embedding for Slot Filling in Natural Language Understanding
- Insights on Neural Representations for End-to-End Speech Recognition
- Causality extraction based on self-attentive BiLSTM-CRF with transferred embeddings
- Beyond Bilinear: Generalized Multimodal Factorized High-Order Pooling for Visual Question Answering
- Deep Spatial Transformation for Pose-Guided Person Image Generation and Animation
- Question Answering over Knowledge Base using Language Model Embeddings
- 3D axial-attention for lung nodule classification
- Artificial intelligence and deep learning algorithms for epigenetic sequence analysis: A review for epigeneticists and AI experts
- A Brief Overview of Unsupervised Neural Speech Representation Learning
- Disinformation, Misinformation, and Fake News in Social Media
- Generating Accurate Assert Statements for Unit Test Cases using Pretrained Transformers
- Two-view correspondence learning using graph neural network with reciprocal neighbor attention
- The Application of Artificial Intelligence to Acoustic Data in Otolaryngology
- Learning Cooperative Multi-Agent Policies With Partial Reward Decoupling
- Tag-less back-translation
- Auto-Embedding Transformer for Interpretable Few-Shot Fault Diagnosis of Rolling Bearings
- Reconstruction of pairwise interactions using energy-based models*
- Spatial-temporal Conv-sequence Learning with Accident Encoding for Traffic Flow Prediction
- Filter-enhanced MLP is All You Need for Sequential Recommendation
- Neural Document Expansion with User Feedback
- Structure-Tags Improve Text Classification for Scholarly Document Quality Prediction
- Novel dual-channel long short-term memory compressed capsule networks for emotion recognition
- A Survey of Orthographic Information in Machine Translation
- Deep Neural Networks and Tabular Data: A Survey
- Future Lens: Anticipating Subsequent Tokens from a Single Hidden State
- Reduce and Reconstruct: ASR for Low-Resource Phonetic Languages
- An LLM-Driven Chatbot in Higher Education for Databases and Information Systems
- SwinNet: Swin Transformer drives edge-aware RGB-D and RGB-T salient object detection
- Creativity in translation: machine translation as a constraint for literary texts
- Predicting Temporal Sets with Deep Neural Networks
- Network On Network for Tabular Data Classification in Real-world Applications
- Heterogeneous Attentions for Solving Pickup and Delivery Problem via Deep Reinforcement Learning
- The Impact of Multiple Parallel Phrase Suggestions on Email Input and Composition Behaviour of Native and Non-Native English Writers
Related