Exploiting Generative AI to Scale up Intelligent Tutoring Systems
2023/01/01 by Jakubův, Jan, Chvalovský, Karel, Goertzel, Zarathustra +6 · 196 citations
Computer Science · #Multimodal Machine Learning Applications #Natural Language Processing Techniques #Topic Modeling
paper · pdf · doi:10.4230/lipics.itp.2023.19
openalex publication_date 2023/01/01 · openalex created_date 2023/07/26 · openalex updated_date 2026/07/30
Abstract
As a present to Mizar on its 50th anniversary, we develop an AI/TP system that automatically proves about 60% of the Mizar theorems in the hammer setting. We also automatically prove 75% of the Mizar theorems when the automated provers are helped by using only the premises used in the human-written Mizar proofs. We describe the methods and large-scale experiments leading to these results. This includes in particular the E and Vampire provers, their ENIGMA and Deepire learning modifications, a number of learning-based premise selection methods, and the incremental loop that interleaves growing a corpus of millions of ATP proofs with training increasingly strong AI/TP systems on them. We also present a selection of Mizar problems that were proved automatically.
Cited by
- A ConvNet for the 2020s
- MetaFormer is Actually What You Need for Vision
- Next word prediction for Urdu language using deep learning models
- Pyramid Vision Transformer: A Versatile Backbone for Dense Prediction without Convolutions
- CrossViT: Cross-Attention Multi-Scale Vision Transformer for Image Classification
- Immersive exposure to simulated visual hallucinations modulates high-level human cognition
- Gender biases and hate speech: Promoters and targets in the Argentinean political context
- Learning skillful medium-range global weather forecasting
- A Survey on Deep Learning for Named Entity Recognition
- A Universal Subhypergraph-Assisted Embedding Framework for Both Homogeneous and Heterogeneous Networks
- S-MGHSTN: Towards An Effective Streaming Traffic Accident Risk Prediction Framework
- High-Resolution Image Synthesis with Latent Diffusion Models
- Skilful precipitation nowcasting using deep generative models of radar
- Swin Transformer: Hierarchical Vision Transformer using Shifted Windows
- Learning defects from aircraft NDT data
- Acoustic source localization by deep-learning attention-based modulation of microphone array data
- How far can you go? Extrapolating values of catalytic activity from known protein landscapes in natural and directed evolution
- ProGen2: Exploring the boundaries of protein language models
- Large Language Models and Generative AI, Oh My!
- Recent Advances in Named Entity Recognition: A Comprehensive Survey and Comparative Study
- Deep Graph Memory Networks for Forgetting-Robust Knowledge Tracing
- Effective and Efficient Multi-View Imputation With Optimal Transport
- Exclusive: the most-cited papers of the twenty-first century
- Vision Transformers for X-ray Diffraction Patterns Analysis
- Large Language Models and the Future of Organization Theory
- The TESCREAL bundle: Eugenics and the promise of utopia through artificial general intelligence
- How to Use Generative AI in Educational Research
- Remembering without (representational) memory: a neuro-computational study on regaining categoricity and compositionality from minimal traces
- The generative era of medical AI
- Methylomes Reveal Recent Evolutionary Changes in Populations of Two Plant Species
- Cyberdelics: Virtual reality hallucinations modulate cognitive-affective processes
- Benchmarking DNA large language models on quadruplexes
- Pixels and Predictions: Potential of GPT-4V in Meteorological Imagery Analysis and Forecast Communication
- Revealing Rubric Relations: Investigating the Interdependence of a Research-Informed and a Machine Learning-Based Rubric in Assessing Student Reasoning in Chemistry
- Multimodal mixing convolutional neural network and transformer for Alzheimer’s disease recognition
- Heterogeneous multivariate time series imputation by transformer model with missing position encoding
- Short-term photovoltaic power forecasting with feature extraction and attention mechanisms
- QPIC: Query-Based Pairwise Human-Object Interaction Detection with Image-Wide Contextual Information
- U-net with ResNet-34 backbone for dual-polarized C-band baltic sea-ice SAR segmentation
- CSwin-PNet: A CNN-Swin Transformer combined pyramid network for breast lesion segmentation in ultrasound images
- Hidformer: Hierarchical dual-tower transformer using multi-scale mergence for long-term time series forecasting
- A deep learning model for multi-modal spatio-temporal irradiance forecast
- Latent diffusion model for conditional reservoir facies generation
- Can we Trust Chatbots for now? Accuracy, reproducibility, traceability; a Case Study on Leonardo da Vinci's Contribution to Astronomy
- Leveraging large language models to predict antibody biological activity against influenza A hemagglutinin
- Sequential Recommendation with Graph Neural Networks
- A Multimodal Deep Learning Approach for Soil Moisture Downscaling Using Remote Sensing and Weather Data
- SwinCrack: Pavement crack detection using convolutional swin-transformer network
- A Survey of Fake News: Fundamental Theories, Detection Methods, and Opportunities
- nnFormer: Volumetric Medical Image Segmentation via a 3D Transformer
- TOPIQ: A Top-Down Approach From Semantics to Distortions for Image Quality Assessment
- Quotation Recommendation and Interpretation Based on Transformation from Queries to Quotations
- Rethinking Semantic Segmentation from a Sequence-to-Sequence Perspective with Transformers
- RGAnomaly: Data reconstruction-based generative adversarial networks for multivariate time series anomaly detection in the Internet of Things
- What Large Language Models Know
- Out of Context: How important is Local Context in Neural Program Repair?
- End-to-End Temporal Action Detection With Transformer
- Fuzzy-ViT: A Deep Neuro-Fuzzy System for Cross-Domain Transfer Learning From Large-Scale General Data to Medical Image
- An Energy-Efficient GeMM-Based Convolution Accelerator With On-the-Fly im2col
- Speech Emotion Recognition via CNN-Transformer and multidimensional attention mechanism
- Diffusion Models in Vision: A Survey
- SimCSE: Simple Contrastive Learning of Sentence Embeddings
- Complex business ecosystem intelligence using AI-powered visual analytics
- Latent-KalmanNet: Learned Kalman Filtering for Tracking From High-Dimensional Signals
- Recognition of European mammals and birds in camera trap images using deep neural networks
- A review of uncertainty quantification in deep learning: Techniques, applications and challenges
- Learning to Prompt for Vision-Language Models
- A Mathematical Investigation of Hallucination and Creativity in GPT Models
- Good for Misconceived Reasons: An Empirical Revisiting on the Need for Visual Context in Multimodal Machine Translation
- E-DSSR: Efficient Dynamic Surgical Scene Reconstruction with Transformer-based Stereoscopic Depth Perception
- Out-of-distribution generalization via composition: A lens through induction heads in Transformers
- CSWin Transformer: A General Vision Transformer Backbone with Cross-Shaped Windows
- Large language models and their applications in bioinformatics
- Image Quality Assessment Using Contrastive Learning
- FocalTransNet: A Hybrid Focal-Enhanced Transformer Network for Medical Image Segmentation
- Mapping the unseen in practice: comparing latent Dirichlet allocation and BERTopic for navigating topic spaces
- learnMSA2: deep protein multiple alignments with large language and hidden Markov models
- On Masked Pre-training and the Marginal Likelihood
- Unifying Large Language Models and Knowledge Graphs: A Roadmap
- Collaborative Forensic Autopsy Documentation and Supervised Report Generation Using a Hybrid Mixed-Reality Environment and Generative AI
- Effectiveness of retrieval augmented generation-based large language models for generating construction safety information
- Audio Mamba: Bidirectional State Space Model for Audio Representation Learning
- Swin Transformer V2: Scaling Up Capacity and Resolution
- Cross-disciplinary perspectives on the potential for artificial intelligence across chemistry
- Adapting a global plant identification model to detect invasive alien plant species in high-resolution road side images
- SqueezeCall: nanopore basecalling using a Squeezeformer network
- Are protein language models the new universal key?
- Latent Space Probing for Adult Content Detection in Video Generative Models
- Video Swin Transformer
- <scp>CerviFormer</scp>: A pap smear‐based cervical cancer classification method using cross‐attention and latent transformer
- ChatGPT for good? On opportunities and challenges of large language models for education
- A Survey on Evaluation of Large Language Models
- BinaryBERT: Pushing the Limit of BERT Quantization
- A Comprehensive Survey of Dynamic Graph Neural Networks: Models, Frameworks, Benchmarks, Experiments and Challenges
- SingSong: Generating musical accompaniments from singing
- Structure-Preserving Transformers for Sequences of SPD Matrices
- Variational Open-Domain Question Answering
- Diffusion tensor estimation with transformer neural networks
- A Mechanistic Interpretation of Arithmetic Reasoning in Language Models using Causal Mediation Analysis
- Data Science in the Big Data Era: Analytics, Intelligence, and Future Challenges
- Multistep traffic forecasting by dynamic graph convolution: Interpretations of real-time spatial correlations
- Human autonomy with AI in the loop
- DeepCodon: A deep learning codon-optimization model to enhance protein expression
- Generating Meaning: Active Inference and the Scope and Limits of Passive AI
- Clinical-grade AI model for molecular subtyping of endometrial cancer: a multi-center cohort study in China
- The Open Science of Deep Learning: Three Case Studies
- Dataset Balancing Can Hurt Model Performance
- Conformer: Convolution-augmented Transformer for Speech Recognition
- Uformer: A General U-Shaped Transformer for Image Restoration
- Time-series forecasting with deep learning: a survey
- Case reports unlocked: Harnessing large language models to advance research on child maltreatment
- Digital authenticity: Towards a research agenda for the AI-driven fifth phase of digitalization in business-to-business marketing
- Human heuristics for AI-generated language are flawed
- What Are We Automating? On the Need for Vision and Expertise When Deploying AI Systems
- The problem of alignment
- <scp>AI</scp> Methods for Antimicrobial Peptides: Progress and Challenges
- SparsePoser: Real-time Full-body Motion Reconstruction from Sparse Data
- Generative AI Solutions to Empower Financial Firms
- Privacy and Security Concerns in Generative AI: A Comprehensive Survey
- Using large language models to extract plant functional traits from unstructured text
- From text to traits: exploring the role of large language models in plant breeding
- A systematic evaluation of Dutch large language models’ surprisal estimates in sentence, paragraph and book reading
- Artificial intelligence in horticulture
- Asymptotic theory of in-context learning by linear attention
- MixingDTA: improved drug–target affinity prediction by extending mixup with guilt-by-association
- Machine learning-driven breakthroughs in water electrolysis and supercapacitors
- An Improved Framework for Scaling Party Positions from Texts with Transformer
- Modelling and design of transcriptional enhancers
- Foundation models in bioinformatics
- Time series predictions in unmonitored sites: a survey of machine learning techniques in water resources
- Protocol paper: From Chaos to Order. Augmenting Manual Article Screening with Sentence Transformers in Management Systematic Reviews
- Beyond the Turing Test: Exploring the implications of generative AI for category construction
- Structure in Deep Reinforcement Learning: A Survey and Open Problems
- GPT (Generative Pre-Trained Transformer)— A Comprehensive Review on Enabling Technologies, Potential Applications, Emerging Challenges, and Future Directions
- Archetypal crop trait dynamics for enhanced retrieval of biophysical parameters from Sentinel-2 MSI
- Theoretical Limitations of Self-Attention in Neural Sequence Models
- Bioeconomy firms and where to find them
- Grounded language acquisition through the eyes and ears of a single child
- Low-cost, autonomous microscopy using deep learning and robotics: A crystal morphology case study
- Generative power of a protein language model trained on multiple sequence alignments
- Advancing Automated Content Analysis for a New Era of Media Effects Research: The Key Role of Transfer Learning
- Facing & mitigating common challenges when working with real-world data: The Data Learning Paradigm
- Can ChatGPT pass Glycobiology?
- AI-textuality: Expanding intertextuality to theorize human-AI interaction with generative artificial intelligence
- Using Emotion Embeddings to Transfer Knowledge Between Emotions, Languages, and Annotation Formats
- How fine can fine-tuning be? Learning efficient language models
- Governance of Generative AI
- How to apply zero‐shot learning to text data in substance use research: An overview and tutorial with media data
- Transformer-based audio-visual multimodal fusion for fine-grained recognition of individual sow nursing behaviour
- EDVR: Video Restoration with Enhanced Deformable Convolutional Networks
- MTKGR: multi-task knowledge graph reasoning for food and ingredient recognition
- CodeBERT: A Pre-Trained Model for Programming and Natural Languages
- <scp>FFM</scp> ‐ <scp>ViT</scp> : an efficient fish species classification method based on deep features and transformers
- Key-value memory in the brain
- Fusing theory-guided machine learning and bio-sensing: considering time in how children learn science from dynamic multimedia
- Large language models can segment narrative events similarly to humans
- GCT: A Granger-Causal Transformer for Multivariate Traffic Analysis in Smart Villages
- When and how to disclose AI use in academic publishing: AMEE Guide No.192
- A systematic literature review of artificial intelligence (AI) in coaching: insights for future research and product development
- UNetFormer: A UNet-like transformer for efficient semantic segmentation of remote sensing urban scene imagery
- Comparison of traditional machine learning and neural network approaches for automated scoring of second language English essays
- Is Genre Enough? A Theory of Genre Signaling as Generative AI Rhetoric
- Leveraging learned representations and multitask learning for lysine methylation site discovery
- Small, open-source text-embedding models as substitutes to OpenAI models for gene analysis
- Computational nanobody design through deep generative modeling and epitope landscape profiling
- GPS: Harnessing data fusion strategies to improve the accuracy of machine learning-based genomic and phenotypic selection
- Flashzoi: An enhanced Borzoi for accelerated genomic analysis
- Symbol ungrounding: what the successes (and failures) of large language models reveal about human cognition
- Autonomy 2.0: The Quest for Economies of Scale
- PIP-Net: Pedestrian Intention Prediction in the Wild
- A large language model-based agent for wayfinding: simulation of spatial perception and memory
- A scoping review of ChatGPT research in accounting and finance
- Using artificial intelligence to document the hidden RNA virosphere
- Language in the Godless Age of AI
- Evaluation of predictions of disordered binding regions in the CAID2 experiment
- SASD: Self-Attention for Small Datasets—A case study in smart villages
- TransHLA: a Hybrid Transformer model for HLA-presented epitope detection
- Transformers and genome language models
- Structural network measures reveal the emergence of heavy-tailed degree distributions in lottery ticket multilayer perceptrons
- Enhancing predictions of protein stability changes induced by single mutations using MSA-based language models
- Security and Privacy Challenges of Large Language Models: A Survey
- Deep learning and generative artificial intelligence in aging research and healthy longevity medicine
- Conducting Qualitative Interviews with AI
- Leveraging mRNA technology for antigen based immuno-oncology therapies
- SCHA-VAE: Hierarchical Context Aggregation for Few-Shot Generation
- HybridDBRpred: improved sequence-based prediction of DNA-binding amino acids using annotations from structured complexes and disordered proteins
- Language model-based B cell receptor sequence embeddings can effectively encode receptor specificity
- Enhanced prediction of vegetation responses to extreme drought using deep learning and Earth observation data
- Adversarial machine learning :
- Optimizing protein tokenization: reduced amino acid alphabets for efficient and accurate protein language models
- Assessing political bias in AI systems: a framework for disentangling viewpoint preferences from epistemic integrity
- From unannotated to annotated
- CamemBERT: a Tasty French Language Model
- OA-CNNs: Omni-Adaptive Sparse CNNs for 3D Semantic Segmentation
- Protein engineering in the deep learning era
- Deep Learning for 3D Point Clouds: A Survey
- Replay and Synthetic Speech Detection with Res2net Architecture
- The Programmer’s Assistant: Conversational Interaction with a Large Language Model for Software Development
- Deep Learning the Functional Renormalization Group
- VGG-TSwinformer: Transformer-based deep learning model for early Alzheimer’s disease prediction
- Opportunities and Challenges in Applying AI to Evolutionary Morphology
- A Multilevel Multimodal Fusion Transformer for Remote Sensing Semantic Segmentation
- TTSR: A Transformer-Based Topography Neural Network for Digital Elevation Model Super-Resolution
- Recent advances on federated learning: A systematic survey
- Guided Attention for Interpretable Motion Captioning
- Vision-Language Models for Vision Tasks: A Survey
- Aggregated Mutual Learning between CNN and Transformer for semi-supervised medical image segmentation
- SPMFormer: Simplified Physical Model-based transformer with cross-space loss for underwater image enhancement
- Adaptive representation-aligned modeling for visual tracking
- Swin Transformer-Based Multiscale Attention Model for Landslide Extraction From Large-Scale Area
- MMD-MLP: LiDAR-Guided Hyperspectral Data Classification Using Local–Global Directional-MLP With Multiresolution Multiscale Representation
- Learning Effective Representations for Person-Job Fit by Feature Fusion
- Discriminative Adversarial Search for Abstractive Summarization
- Artificial Neural Networks for Neuroscientists: A Primer
- Weakly-supervised Fingerspelling Recognition in British Sign Language\n Videos
- DeltaConv
- Literaturwissenschaft und Informatik
- An Attention-Based Deep Learning Approach for Sleep Stage Classification With Single-Channel EEG
- Shortcut learning in deep neural networks
- Tokens-to-Token ViT: Training Vision Transformers from Scratch on ImageNet
- Raster‐to‐Graph: Floorplan Recognition via Autoregressive Graph Prediction with an Attention Transformer
- A time-series neural network for pig feeding behavior recognition and dangerous detection from videos
- Deep learning in multiple animal tracking: A survey
- Exploiting BERT for End-to-End Aspect-based Sentiment Analysis
- Emerging Properties in Self-Supervised Vision Transformers
- BNAI, NO-TOKEN, and MIND-UNITY: Pillars of a Systemic Revolution in Artificial Intelligence
- Scalable Coupling of Deep Learning with Logical Reasoning
- A Practical Deep Learning-Based Acoustic Side Channel Attack on Keyboards
- Point Transformer
- Goal-conditioned dual-action imitation learning for dexterous dual-arm robot manipulation
- Motion Retargetting based on Dilated Convolutions and Skeleton‐specific Loss Functions
- Medical Image Segmentation Review: The Success of U-Net
- Madness, Cannibalism, and Traditional Fiction between Lu Xun and Mo Yan
- Foundation models for electrocardiogram interpretation: clinical implications
- Arabic Automatic Speech Recognition: Challenges and Progress
- Transformer-based Self-supervised Multimodal Representation Learning for Wearable Emotion Recognition
- Informer: Beyond Efficient Transformer for Long Sequence Time-Series Forecasting
- Development and Validation of a Natural Language Processing Algorithm to Pseudonymize Documents in the Context of a Clinical Data Warehouse
- EVA-02: A visual representation for neon genesis
- Practical and ethical challenges of large language models in education: A systematic scoping review
- Thinking Fast and Slow: Efficient Text-to-Visual Retrieval with\n Transformers
- Beyond Individual Input for Deep Anomaly Detection on Tabular Data
- Language Conditioned Spatial Relation Reasoning for 3D Object Grounding
- Efficient Diffusion Model for Image Restoration by Residual Shifting
- DiffusionAD: Norm-Guided One-Step Denoising Diffusion for Anomaly Detection
- An algorithmic framework for the optimization of deep neural networks architectures and hyperparameters
- One-shot Unsupervised Domain Adaptation with Personalized Diffusion Models
- CALYPSO: LLMs as Dungeon Master's Assistants
- Balsa: Learning a Query Optimizer Without Expert Demonstrations
- Experimental Standards for Deep Learning in Natural Language Processing Research
- PVT v2: Improved baselines with pyramid vision transformer
- DiffuRec: A Diffusion Model for Sequential Recommendation
- Multimodal attention-based deep learning for Alzheimer’s disease diagnosis
- An Attention-based Graph Neural Network for Heterogeneous Structural Learning
- The benefits and dangers of anthropomorphic conversational agents
- An antimicrobial drug recommender system using MALDI-TOF MS and dual-branch neural networks
- Predicting stroke outcome: A case for multimodal deep learning methods with tabular and CT Perfusion data
- From the Token to the Review: A Hierarchical Multimodal approach to\n Opinion Mining
- Are Transformers Effective for Time Series Forecasting?
- PICK: Processing Key Information Extraction from Documents using Improved Graph Learning-Convolutional Networks
- Delta Keyword Transformer: Bringing Transformers to the Edge through Dynamically Pruned Multi-Head Self-Attention
- Hybrid semantics-based vulnerability detection incorporating a Temporal Convolutional Network and Self-attention Mechanism
- VAERHNN: Voting-averaged ensemble regression and hybrid neural network to investigate potent leads against colorectal cancer
- Navigating the Complexity of Generative AI Adoption in Software Engineering
- Continuous Entity Reasoning for Multi-Turn Medical Dialog Generation
- Plot and Rework: Modeling Storylines for Visual Storytelling
- Introducing AIRSim: An Innovative AI-Driven Feedback Generation Tool for Supporting Student Learning
- Critique of impure reason: Unveiling the reasoning behaviour of medical large language models
Related