Kevin Jamieson
- Non-stochastic Best Arm Identification and Hyperparameter Optimization
2015/02/27 by Kevin Jamieson, Jamieson, Kevin, Ameet Talwalkar +1 · 19 citations
Computer Science · Decision Sciences · #Advanced Bandit Algorithms Research #FOS: Computer and information sciences #Machine Learning (cs.LG) #Machine Learning (stat.ML) #Machine Learning and Algorithms #Machine Learning and Data Classification
- lil' UCB : An Optimal Exploration Algorithm for Multi-Armed Bandits
2013/12/27 by Kevin Jamieson, Matthew Malloy, Jamieson, Kevin +5 · 12 citations
Computer Science · Decision Sciences · #Advanced Bandit Algorithms Research #Auction Theory and Applications #FOS: Computer and information sciences #Machine Learning (cs.LG) #Machine Learning (stat.ML) #Reinforcement Learning in Robotics
- Active Ranking using Pairwise Comparisons
2011/09/16 by Kevin Jamieson, Jamieson, Kevin G., Robert D. Nowak +1 · 11 citations
Computer Science · Economics, Econometrics and Finance · #Data Management and Algorithms #FOS: Computer and information sciences #Game Theory and Voting Systems #Information Theory (cs.IT) #Machine Learning (cs.LG) #Machine Learning (stat.ML)
- Non-Asymptotic Gap-Dependent Regret Bounds for Tabular MDPs
2019/05/09 by Max Simchowitz, Simchowitz, Max, Kevin Jamieson +1 · 11 citations
Computer Science · Decision Sciences · Engineering · #Advanced Bandit Algorithms Research #FOS: Computer and information sciences #FOS: Mathematics #Machine Learning (cs.LG) #Machine Learning (stat.ML) #Optimization and Control (math.OC) #Reinforcement Learning in Robotics #Smart Grid Energy Management #Statistics Theory (math.ST)
- Query Complexity of Derivative-Free Optimization
2012/09/11 by Kevin Jamieson, Jamieson, Kevin G., Robert D. Nowak +3 · 4 citations
Computer Science · #Complexity and Algorithms in Graphs #Computability, Logic, AI Algorithms #FOS: Computer and information sciences #Machine Learning (cs.LG) #Machine Learning (stat.ML) #Machine Learning and Algorithms
- First-Order Regret in Reinforcement Learning with Linear Function Approximation: A Robust Estimation Approach
2021/12/07 by Andrew Wagenmaker, Yifang Chen, Wagenmaker, Andrew +7 · 4 citations
Computer Science · Decision Sciences · Engineering · #Advanced Bandit Algorithms Research #FOS: Computer and information sciences #Machine Learning (cs.LG) #Machine Learning (stat.ML) #Reinforcement Learning in Robotics #Smart Grid Energy Management
- Humor in AI: Massive Scale Crowd-Sourced Preferences and Benchmarks for Cartoon Captioning
2024/06/15 by Jifan Zhang, Zhang, Jifan, Lalit Jain +21 · 2 voices · 5 citations
Arts and Humanities · Psychology · #Artificial Intelligence (cs.AI) #Computation and Language (cs.CL) #FOS: Computer and information sciences #Humor Studies and Applications #Language, Metaphor, and Cognition #Machine Learning (cs.LG) #Subtitles and Audiovisual Media
- Optimal Exploration for Model-Based RL in Nonlinear Systems
2023/06/15 by Andrew Wagenmaker, Guanya Shi, Wagenmaker, Andrew +3 · 4 citations
Biochemistry, Genetics and Molecular Biology · Engineering · Computer Science · #Receptor Mechanisms and Signaling #Advanced Control Systems Optimization #Machine Learning and Algorithms
- The True Sample Complexity of Identifying Good Arms
2019/06/15 by Julian Katz-Samuels, Kevin Jamieson, Katz-Samuels, Julian +1 · 3 citations
Decision Sciences · Computer Science · #Advanced Bandit Algorithms Research #Machine Learning and Algorithms #Gaussian Processes and Bayesian Inference
- The Simulator: Understanding Adaptive Sampling in the\n Moderate-Confidence Regime
2017/02/16 by Max Simchowitz, Simchowitz, Max, Kevin Jamieson +3 · 3 citations
Computer Science · Decision Sciences · #Advanced Bandit Algorithms Research #FOS: Computer and information sciences #Machine Learning (cs.LG) #Machine Learning (stat.ML) #Machine Learning and Algorithms #Optimization and Search Problems
- Finite Sample Prediction and Recovery Bounds for Ordinal Embedding
2016/06/22 by Lalit Jain, Jain, Lalit, Kevin Jamieson +3 · 2 citations
Engineering · Physics and Astronomy · Computer Science · #Sparse and Compressive Sensing Techniques #Model Reduction and Neural Networks #Face and Expression Recognition
- Task-Optimal Exploration in Linear Dynamical Systems
2021/02/10 by Andrew Wagenmaker, Max Simchowitz, Wagenmaker, Andrew +3 · 4 citations
Decision Sciences · Computer Science · Engineering · #Advanced Bandit Algorithms Research #Reinforcement Learning in Robotics #Advanced Control Systems Optimization
- CLIPLoss and Norm-Based Data Selection Methods for Multimodal Contrastive Learning
2024/05/29 by Yiping Wang, Yifang Chen, Wang, Yiping +11 · 1 voice · 4 citations
Computer Science · #Computer Vision and Pattern Recognition (cs.CV) #FOS: Computer and information sciences #Machine Learning (cs.LG) #cs.CV #cs.LG
- Best Arm Identification with Safety Constraints
2021/11/23 by Zhenlin Wang, Andrew Wagenmaker, Wang, Zhenlin +3 · 3 citations
Computer Science · Decision Sciences · #Advanced Bandit Algorithms Research #FOS: Computer and information sciences #Machine Learning (cs.LG) #Machine Learning (stat.ML) #Machine Learning and Algorithms #Reinforcement Learning in Robotics
- An Experimental Design Framework for Label-Efficient Supervised Finetuning of Large Language Models
2024/01/12 by Gantavya Bhatt, Bhatt, Gantavya, Yifang Chen +21 · 4 citations
Computer Science · #Artificial Intelligence (cs.AI) #Computation and Language (cs.CL) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Machine Learning and Algorithms #Natural Language Processing Techniques #Topic Modeling
- Instance-Dependent Near-Optimal Policy Identification in Linear MDPs via Online Experiment Design
2022/07/06 by Andrew Wagenmaker, Kevin Jamieson, Wagenmaker, Andrew +1 · 2 citations
Computer Science · Decision Sciences · #Advanced Bandit Algorithms Research #FOS: Computer and information sciences #Machine Learning (cs.LG) #Machine Learning (stat.ML) #Machine Learning and Algorithms #Machine Learning and Data Classification
- A framework for Multi-A(rmed)/B(andit) testing with online FDR control
2017/06/16 by Fanny Yang, Aaditya Ramdas, Yang, Fanny +5 · 1 citation
Computer Science · #VLSI and Analog Circuit Testing
- Overcoming the Sim-to-Real Gap: Leveraging Simulation to Learn to Explore for Real-World RL
2024/10/26 by Andrew Wagenmaker, Wagenmaker, Andrew, Kevin Huang +9 · 3 citations
Decision Sciences · Computer Science · #Simulation Techniques and Applications #Model-Driven Software Engineering Techniques
- Mosaic: A Sample-Based Database System for Open World Query Processing
2019/12/17 by Laurel Orr, Samuel Ainsworth, Orr, Laurel +9 · 1 citation
Computer Science · Decision Sciences · #Data Management and Algorithms #Advanced Database Systems and Queries #Data Quality and Management
- Improved Corruption Robust Algorithms for Episodic Reinforcement\n Learning
2021/02/13 by Yifang Chen, Simon S. Du, Chen, Yifang +3 · 1 citation
Computer Science · Decision Sciences · Engineering · #Advanced Bandit Algorithms Research #Artificial Intelligence (cs.AI) #Auction Theory and Applications #FOS: Computer and information sciences #Machine Learning (cs.LG) #Reinforcement Learning in Robotics #Smart Grid Energy Management
- Beyond No Regret: Instance-Dependent PAC Reinforcement Learning
2021/08/05 by Andrew Wagenmaker, Wagenmaker, Andrew, Max Simchowitz +3 · 1 citation
Computer Science · Decision Sciences · #Advanced Bandit Algorithms Research #Adversarial Robustness in Machine Learning #FOS: Computer and information sciences #Machine Learning (cs.LG) #Machine Learning (stat.ML) #Reinforcement Learning in Robotics
- Improved Active Multi-Task Representation Learning via Lasso
2023/06/05 by Yiping Wang, Wang, Yiping, Yifang Chen +5 · 1 citation
Computer Science · Engineering · Decision Sciences · #Domain Adaptation and Few-Shot Learning #Sparse and Compressive Sensing Techniques #Advanced Bandit Algorithms Research
- Nearly Minimax Optimal Submodular Maximization with Bandit Feedback
2023/10/27 by Artin Tajdini, Lalit Jain, Tajdini, Artin +3 · 1 citation
Decision Sciences · Computer Science · #Advanced Bandit Algorithms Research #Machine Learning and Algorithms #Distributed Sensor Networks and Detection Algorithms
- Variance Alignment Score: A Simple But Tough-to-Beat Data Selection Method for Multimodal Contrastive Learning
2024/02/03 by Yiping Wang, Yifang Chen, Wang, Yiping +7 · 1 citation
Computer Science · #Artificial Intelligence (cs.AI) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Speech and dialogue systems #Text and Document Classification Technologies
- Improved Regret Bounds for Linear Bandits with Heavy-Tailed Rewards
2025/06/05 by Artin Tajdini, Jonathan Scarlett, Tajdini, Artin +3 · 2 citations
Computer Science · Decision Sciences · #Advanced Bandit Algorithms Research #Age of Information Optimization #FOS: Computer and information sciences #Information Theory (cs.IT) #Machine Learning (cs.LG) #Machine Learning (stat.ML) #Stochastic Gradient Optimization Techniques
- Efficient Near-Optimal Algorithm for Online Shortest Paths in Directed Acyclic Graphs with Bandit Feedback Against Adaptive Adversaries
2025/04/01 by Arnab Maiti, Maiti, Arnab, Zhiyuan Fan +7 · 3 citations
Computer Science · Decision Sciences · #Advanced Bandit Algorithms Research #Computer Science and Game Theory (cs.GT) #FOS: Computer and information sciences #Game Theory and Applications #Machine Learning (cs.LG) #Optimization and Search Problems