Can large language models be a cardinality estimator? An empirical study
2026/03/31 by Liangzu Liu, Yinjun Wu, Yiyan Wang +8
paper · doi:10.1007/s00778-026-00975-7
Citations
- Register Always Matters: Analysis of LLM Pretraining Data Through the Lens of Language Variation
- The Llama 3 Herd of Models
- Regurgitative Training: The Value of Real Data in Training Large Language Models
- PRICE: A Pretrained Model for Cross-Database Cardinality Estimation
- Does Fine-Tuning LLMs on New Knowledge Encourage Hallucinations?
- DeepSeek-V2: A Strong, Economical, and Efficient Mixture-of-Experts Language Model
- Get more for less: Principled Data Selection for Warming Up Fine-Tuning in LLMs
- LLM-R2: A Large Language Model Enhanced Rule-based Rewrite System for Boosting Query Efficiency
- LlamaFactory: Unified Efficient Fine-Tuning of 100+ Language Models
- CodeS: Towards Building Open-source Language Models for Text-to-SQL
- QuRating: Selecting High-Quality Data for Training Language Models
- D-Bot: Database Diagnosis System using Large Language Models
- A Survey on Hallucination in Large Language Models: Principles, Taxonomy, Challenges, and Open Questions
- The Troubling Emergence of Hallucination in Large Language Models -- An Extensive Definition, Quantification, and Prescriptive Remediations
- Automatically Correcting Large Language Models: Surveying the landscape of diverse self-correction strategies
- Few-shot Fine-tuning vs. In-context Learning: A Fair Comparison and Evaluation
- Small Models are Valuable Plug-ins for Large Language Models
- Visual Instruction Tuning
- Self-Refine: Iterative Refinement with Self-Feedback
- Reflexion: Language Agents with Verbal Reinforcement Learning
- SelfCheckGPT: Zero-Resource Black-Box Hallucination Detection for Generative Large Language Models
- Prediction-Powered Inference
- Detect, Distill and Update: Learned DB Systems Facing Out of Distribution Data
- Training Compute-Optimal Large Language Models
- Cardinality Estimation in DBMS: A Comprehensive Benchmark Evaluation
- Evaluating Large Language Models Trained on Code
- Flow-Loss: Learning Cardinality Estimates That Matter
- NeuroCard: One Cardinality Estimator for All Tables
- On Faithfulness and Factuality in Abstractive Summarization
- NN-based Transformation of Any SQL Cardinality Estimator for Handling DISTINCT, AND, OR and NOT
- DeepDB: Learn from Data, not from Queries!
- Learned Cardinalities: Estimating Correlated Joins with Deep Learning