Akshay Krishnamurthy
- Transformers Learn Shortcuts to Automata
2022/10/19 by Bingbin Liu, Liu, Bingbin, Jordan T. Ash +7 · 3 voices · 39 citations
Computer Science · #Algorithms and Data Compression #Machine Learning and Algorithms #Topic Modeling #cs.FL #cs.LG #stat.ML
- Go for a Walk and Arrive at the Answer: Reasoning Over Paths in Knowledge Bases using Reinforcement Learning
2017/11/15 by Rajarshi Das, Shehzaad Dhuliawala, Das, Rajarshi +13 · 35 citations
Computer Science · #Topic Modeling #Natural Language Processing Techniques #Advanced Graph Neural Networks
- Contextual Decision Processes with Low Bellman Rank are PAC-Learnable
2016/10/29 by Nan Jiang, Akshay Krishnamurthy, Jiang, Nan +7 · 25 citations
Computer Science · Engineering · #Bayesian Modeling and Causal Inference #FOS: Computer and information sciences #Fault Detection and Control Systems #Machine Learning (cs.LG) #Machine Learning (stat.ML) #Neural Networks and Applications
- Self-Improvement in Language Models: The Sharpening Mechanism
2024/12/02 by Audrey Huang, Adam Block, Huang, Audrey +13 · 2 voices · 27 citations
Computer Science · #Natural Language Processing Techniques #Topic Modeling #cs.AI #cs.CL #cs.LG #stat.ML
- Provably efficient RL with Rich Observations via Latent State Decoding
2019/01/25 by Simon S. Du, Akshay Krishnamurthy, Du, Simon S. +9 · 17 citations
Computer Science · #Machine Learning and Algorithms #Data Stream Mining Techniques #Adversarial Robustness in Machine Learning
- Can large language models explore in-context?
2024/03/22 by Akshay Krishnamurthy, Krishnamurthy, Akshay, Keegan Harris +7 · 2 voices · 15 citations
Computer Science · #Natural Language Processing Techniques
- Learning to Search Better Than Your Teacher
2015/02/08 by Kai-Wei Chang, Akshay Krishnamurthy, Chang, Kai-Wei +7 · 12 citations
Decision Sciences · Computer Science · #Advanced Bandit Algorithms Research #Machine Learning and Algorithms #Reinforcement Learning in Robotics
- Exploratory Preference Optimization: Harnessing Implicit Q*-Approximation for Sample-Efficient RLHF
2024/05/31 by Tengyang Xie, Dylan J. Foster, Xie, Tengyang +10 · 1 voice · 19 citations
Computer Science · Mathematics · #Advanced Data Compression Techniques #Artificial Intelligence (cs.AI) #Computation and Language (cs.CL) #Data Management and Algorithms #FOS: Computer and information sciences #Face and Expression Recognition #Machine Learning (cs.LG) #Machine Learning (stat.ML) #cs.AI #cs.CL #cs.LG #stat.ML
- Off-policy evaluation for slate recommendation
2016/05/16 by Adith Swaminathan, Swaminathan, Adith, Akshay Krishnamurthy +11 · 8 citations
Computer Science · Decision Sciences · #Advanced Bandit Algorithms Research #Artificial Intelligence (cs.AI) #Auction Theory and Applications #FOS: Computer and information sciences #Machine Learning (cs.LG) #Machine Learning (stat.ML) #Optimization and Search Problems
- Hybrid RL: Using Both Offline and Online Data Can Make RL Efficient
2022/10/13 by Yuda Song, Song, Yuda, Yifei Zhou +9 · 13 citations
Computer Science · Decision Sciences · #Advanced Bandit Algorithms Research #FOS: Computer and information sciences #Machine Learning (cs.LG) #Mobile Crowdsensing and Crowdsourcing #Reinforcement Learning in Robotics
- Is Best-of-N the Best of Them? Coverage, Scaling, and Optimality in Inference-Time Alignment
2025/03/27 by Audrey Huang, Huang, Audrey, Adam Block +9 · 2 voices · 21 citations
#cs.AI #cs.LG #stat.ML
- Exposing Attention Glitches with Flip-Flop Language Modeling
2023/06/01 by Bingbin Liu, Liu, Bingbin, Jordan T. Ash +7 · 1 voice · 9 citations
Computer Science · Engineering · #Computation and Language (cs.CL) #FOS: Computer and information sciences #Ferroelectric and Negative Capacitance Devices #Machine Learning (cs.LG) #Natural Language Processing Techniques #Topic Modeling
- Next-Latent Prediction Transformers Learn Compact World Models
2025/11/08 by Jayden Teoh, J. J. Teoh, Manan Tomar +18 · 4 voices · 3 citations
Computer Science · #cs.LG
- Understanding Contrastive Learning Requires Incorporating Inductive Biases
2022/02/28 by Nikunj Saunshi, Saunshi, Nikunj, Jordan T. Ash +13 · 6 citations
Computer Science · #Artificial Intelligence (cs.AI) #Domain Adaptation and Few-Shot Learning #FOS: Computer and information sciences #Machine Learning (cs.LG) #Machine Learning and Algorithms #Machine Learning and Data Classification
- Sample-Efficient Reinforcement Learning of Undercomplete POMDPs
2020/06/22 by Chi Jin, Jin, Chi, Sham M. Kakade +5 · 5 citations
Computer Science · Engineering · #Artificial Intelligence (cs.AI) #Blind Source Separation Techniques #Elevator Systems and Control #FOS: Computer and information sciences #FOS: Mathematics #Machine Learning (cs.LG) #Machine Learning (stat.ML) #Neural Networks and Applications #Optimization and Control (math.OC)
- Correcting the Mythos of KL-Regularization: Direct Alignment without Overoptimization via Chi-Squared Preference Optimization
2024/07/18 by Audrey Huang, Wenhao Zhan, Huang, Audrey +11 · 1 voice · 9 citations
Computer Science · Engineering · #Advanced Multi-Objective Optimization Algorithms #Artificial Intelligence (cs.AI) #Computation and Language (cs.CL) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Neural Networks and Applications #Topology Optimization in Engineering #cs.AI #cs.CL #cs.LG
- Contextual Bandits with Continuous Actions: Smoothing, Zooming, and Adapting
2019/02/05 by Akshay Krishnamurthy, John Langford, Krishnamurthy, Akshay +5 · 5 citations
Computer Science · Decision Sciences · #Advanced Bandit Algorithms Research #Data Stream Mining Techniques #FOS: Computer and information sciences #Machine Learning (cs.LG) #Machine Learning (stat.ML) #Reinforcement Learning in Robotics
- Efficient First-Order Contextual Bandits: Prediction, Allocation, and Triangular Discrimination
2021/07/05 by Dylan J. Foster, Foster, Dylan J., Akshay Krishnamurthy +1 · 5 citations
Computer Science · Decision Sciences · #Advanced Bandit Algorithms Research #FOS: Computer and information sciences #FOS: Mathematics #Machine Learning (cs.LG) #Machine Learning (stat.ML) #Machine Learning and Algorithms #Machine Learning and Data Classification #Statistics Theory (math.ST)
- Gone Fishing: Neural Active Learning with Fisher Embeddings
2021/06/17 by Jordan T. Ash, Ash, Jordan T., Surbhi Goel +5 · 5 citations
Computer Science · #Adversarial Robustness in Machine Learning #FOS: Computer and information sciences #Machine Learning (cs.LG) #Machine Learning and Algorithms #Machine Learning and Data Classification
- Guaranteed Discovery of Control-Endogenous Latent States with Multi-Step Inverse Models
2022/07/17 by Alex Lamb, Riashat Islam, Lamb, Alex +17 · 5 citations
Computer Science · #Time Series Analysis and Forecasting #Reinforcement Learning in Robotics #AI-based Problem Solving and Planning
- Streaming Active Learning with Deep Neural Networks
2023/03/05 by Akanksha Saran, Saran, Akanksha, Safoora Yousefi +7 · 6 citations
Computer Science · #Machine Learning and Algorithms #Data Stream Mining Techniques #Machine Learning and Data Classification
- Optimism in Reinforcement Learning with Generalized Linear Function Approximation
2019/12/09 by Yining Wang, Ruosong Wang, Wang, Yining +5 · 4 citations
Computer Science · Decision Sciences · Engineering · #Advanced Bandit Algorithms Research #FOS: Computer and information sciences #Machine Learning (cs.LG) #Machine Learning (stat.ML) #Reinforcement Learning in Robotics #Smart Grid Energy Management
- Model-free Representation Learning and Exploration in Low-rank MDPs
2021/02/14 by Aditya Modi, Jing‐Lin Chen, Modi, Aditya +7 · 7 citations
Computer Science · Decision Sciences · Physics and Astronomy · #Reinforcement Learning in Robotics #Advanced Bandit Algorithms Research #Model Reduction and Neural Networks
- Computational-Statistical Tradeoffs at the Next-Token Prediction Barrier: Autoregressive and Imitation Learning under Misspecification
2025/02/18 by Dhruv Rohatgi, Rohatgi, Dhruv, Adam Block +7 · 1 voice · 7 citations
Computer Science · #Natural Language Processing Techniques
- PAC Reinforcement Learning with Rich Observations
2016/02/08 by Akshay Krishnamurthy, Krishnamurthy, Akshay, Alekh Agarwal +3 · 2 citations
Computer Science · Decision Sciences · #Advanced Bandit Algorithms Research #Data Stream Mining Techniques #FOS: Computer and information sciences #Machine Learning (cs.LG) #Machine Learning (stat.ML) #Reinforcement Learning in Robotics
- Efficient and Optimal Algorithms for Contextual Dueling Bandits under\n Realizability
2021/11/24 by Aadirupa Saha, Akshay Krishnamurthy, Saha, Aadirupa +1 · 2 citations
Computer Science · Decision Sciences · #Advanced Bandit Algorithms Research #Auction Theory and Applications #FOS: Computer and information sciences #Machine Learning (cs.LG) #Machine Learning and Algorithms #Optimization and Search Problems #Stochastic Gradient Optimization Techniques
- On Oracle-Efficient PAC RL with Rich Observations
2018/03/01 by Christoph Dann, Dann, Christoph, Nan Jiang +9 · 1 citation
Computer Science · Decision Sciences · #Auction Theory and Applications #Evolutionary Algorithms and Applications #FOS: Computer and information sciences #Machine Learning (cs.LG) #Machine Learning (stat.ML) #Reinforcement Learning in Robotics
- Robust Dynamic Assortment Optimization in the Presence of Outlier Customers
2019/10/09 by Xi Chen, Chen, Xi, Akshay Krishnamurthy +3 · 2 citations
Decision Sciences · Computer Science · Business, Management and Accounting · #Advanced Bandit Algorithms Research #Optimization and Search Problems #Supply Chain and Inventory Management
- Butterfly Effects of SGD Noise: Error Amplification in Behavior Cloning and Autoregression
2023/10/17 by Adam Block, Dylan J. Foster, Block, Adam +7 · 3 citations
Computer Science · #FOS: Computer and information sciences #FOS: Mathematics #Machine Learning (cs.LG) #Machine Learning (stat.ML) #Optimization and Control (math.OC) #Reinforcement Learning in Robotics
- Kinematic State Abstraction and Provably Efficient Rich-Observation Reinforcement Learning
2019/11/13 by Dipendra Misra, Mikael Henaff, Misra, Dipendra +5 · 1 citation
Computer Science · Engineering · #Anomaly Detection Techniques and Applications #FOS: Computer and information sciences #Human Pose and Action Recognition #Machine Learning (cs.LG) #Machine Learning (stat.ML) #Robot Manipulation and Learning
- Semiparametric Contextual Bandits
2018/03/12 by Akshay Krishnamurthy, Krishnamurthy, Akshay, Zhiwei Steven Wu +3 · 2 citations
Decision Sciences · Computer Science · Engineering · #Advanced Bandit Algorithms Research #Reinforcement Learning in Robotics #Smart Grid Energy Management
- Provable RL with Exogenous Distractors via Multistep Inverse Dynamics
2021/10/17 by Yonathan Efroni, Efroni, Yonathan, Dipendra Misra +7 · 1 citation
Computer Science · Decision Sciences · #Advanced Bandit Algorithms Research #Data Stream Mining Techniques #FOS: Computer and information sciences #Machine Learning (cs.LG) #Reinforcement Learning in Robotics
- Oracle-Efficient Pessimism: Offline Policy Optimization in Contextual Bandits
2023/06/13 by Lequn Wang, Wang, Lequn, Akshay Krishnamurthy +3 · 1 citation
Computer Science · Decision Sciences · #Advanced Bandit Algorithms Research #FOS: Computer and information sciences #Machine Learning (cs.LG) #Machine Learning and Algorithms #Machine Learning and Data Classification
- Rich-Observation Reinforcement Learning with Continuous Latent Dynamics
2024/05/29 by Yuda Song, Song, Yuda, Lili Wu +5 · 1 citation
Computer Science · #Data Stream Mining Techniques #Domain Adaptation and Few-Shot Learning #FOS: Computer and information sciences #Machine Learning (cs.LG) #Machine Learning and Data Classification
- Wait, Wait, Wait... Why Do Reasoning Models Loop?
2025/12/15 by Charilaos Pipis, Pipis, Charilaos, Garg, Shivam +8 · 1 citation
Computer Science · Psychology · #Intelligent Tutoring Systems and Adaptive Learning #Topic Modeling #Visual and Cognitive Learning Processes