vix.ing · top · new · best · stats · spec

Akshay Krishnamurthy

  1. Transformers Learn Shortcuts to Automata
    2022/10/19 by Bingbin Liu, Jordan T. Ash, Liu, Bingbin +7 · 3 voices · 37 citations
    Computer Science · #Algorithms and Data Compression #Machine Learning and Algorithms #Topic Modeling #cs.FL #cs.LG #stat.ML
  2. Go for a Walk and Arrive at the Answer: Reasoning Over Paths in Knowledge Bases using Reinforcement Learning
    2017/11/15 by Rajarshi Das, Das, Rajarshi, Shehzaad Dhuliawala +13 · 33 citations
    Computer Science · #Topic Modeling #Natural Language Processing Techniques #Advanced Graph Neural Networks
  3. Contextual Decision Processes with Low Bellman Rank are PAC-Learnable
    2016/10/29 by Nan Jiang, Akshay Krishnamurthy, Jiang, Nan +7 · 24 citations
    Computer Science · Engineering · #Bayesian Modeling and Causal Inference #FOS: Computer and information sciences #Fault Detection and Control Systems #Machine Learning (cs.LG) #Machine Learning (stat.ML) #Neural Networks and Applications
  4. Self-Improvement in Language Models: The Sharpening Mechanism
    2024/12/02 by Audrey Huang, Huang, Audrey, Adam Block +13 · 2 voices · 26 citations
    Computer Science · #Natural Language Processing Techniques #Topic Modeling #cs.AI #cs.CL #cs.LG #stat.ML
  5. Can large language models explore in-context?
    2024/03/22 by Akshay Krishnamurthy, Keegan Harris, Krishnamurthy, Akshay +7 · 2 voices · 15 citations
    Computer Science · #Natural Language Processing Techniques
  6. Provably efficient RL with Rich Observations via Latent State Decoding
    2019/01/25 by Simon S. Du, Du, Simon S., Akshay Krishnamurthy +9 · 16 citations
    Computer Science · #Machine Learning and Algorithms #Data Stream Mining Techniques #Adversarial Robustness in Machine Learning
  7. Learning to Search Better Than Your Teacher
    2015/02/08 by Kai-Wei Chang, Akshay Krishnamurthy, Chang, Kai-Wei +7 · 10 citations
    Decision Sciences · Computer Science · #Advanced Bandit Algorithms Research #Machine Learning and Algorithms #Reinforcement Learning in Robotics
  8. Exploratory Preference Optimization: Harnessing Implicit Q*-Approximation for Sample-Efficient RLHF
    2024/05/31 by Tengyang Xie, Xie, Tengyang, Dylan J. Foster +10 · 1 voice · 19 citations
    Computer Science · Mathematics · #Advanced Data Compression Techniques #Artificial Intelligence (cs.AI) #Computation and Language (cs.CL) #Data Management and Algorithms #FOS: Computer and information sciences #Face and Expression Recognition #Machine Learning (cs.LG) #Machine Learning (stat.ML) #cs.AI #cs.CL #cs.LG #stat.ML
  9. Hybrid RL: Using Both Offline and Online Data Can Make RL Efficient
    2022/10/13 by Yuda Song, Song, Yuda, Yifei Zhou +9 · 11 citations
    Computer Science · Decision Sciences · #Advanced Bandit Algorithms Research #FOS: Computer and information sciences #Machine Learning (cs.LG) #Mobile Crowdsensing and Crowdsourcing #Reinforcement Learning in Robotics
  10. Is Best-of-N the Best of Them? Coverage, Scaling, and Optimality in Inference-Time Alignment
    2025/03/27 by Audrey Huang, Huang, Audrey, Adam Block +9 · 2 voices · 20 citations
    #cs.AI #cs.LG #stat.ML
  11. Off-policy evaluation for slate recommendation
    2016/05/16 by Adith Swaminathan, Swaminathan, Adith, Akshay Krishnamurthy +11 · 6 citations
    Computer Science · Decision Sciences · #Advanced Bandit Algorithms Research #Artificial Intelligence (cs.AI) #Auction Theory and Applications #FOS: Computer and information sciences #Machine Learning (cs.LG) #Machine Learning (stat.ML) #Optimization and Search Problems
  12. Exposing Attention Glitches with Flip-Flop Language Modeling
    2023/06/01 by Bingbin Liu, Jordan T. Ash, Liu, Bingbin +7 · 1 voice · 8 citations
    Computer Science · Engineering · #Computation and Language (cs.CL) #FOS: Computer and information sciences #Ferroelectric and Negative Capacitance Devices #Machine Learning (cs.LG) #Natural Language Processing Techniques #Topic Modeling
  13. Next-Latent Prediction Transformers Learn Compact World Models
    2025/11/08 by Jayden Teoh, J. J. Teoh, Teoh, Jayden +18 · 4 voices · 3 citations
    Computer Science · #cs.LG
  14. Understanding Contrastive Learning Requires Incorporating Inductive Biases
    2022/02/28 by Nikunj Saunshi, Jordan T. Ash, Saunshi, Nikunj +13 · 6 citations
    Computer Science · #Artificial Intelligence (cs.AI) #Domain Adaptation and Few-Shot Learning #FOS: Computer and information sciences #Machine Learning (cs.LG) #Machine Learning and Algorithms #Machine Learning and Data Classification
  15. Sample-Efficient Reinforcement Learning of Undercomplete POMDPs
    2020/06/22 by Chi Jin, Jin, Chi, Sham M. Kakade +5 · 5 citations
    Computer Science · Engineering · #Artificial Intelligence (cs.AI) #Blind Source Separation Techniques #Elevator Systems and Control #FOS: Computer and information sciences #FOS: Mathematics #Machine Learning (cs.LG) #Machine Learning (stat.ML) #Neural Networks and Applications #Optimization and Control (math.OC)
  16. Correcting the Mythos of KL-Regularization: Direct Alignment without Overoptimization via Chi-Squared Preference Optimization
    2024/07/18 by Audrey Huang, Huang, Audrey, Wenhao Zhan +11 · 1 voice · 9 citations
    Computer Science · Engineering · #Advanced Multi-Objective Optimization Algorithms #Artificial Intelligence (cs.AI) #Computation and Language (cs.CL) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Neural Networks and Applications #Topology Optimization in Engineering #cs.AI #cs.CL #cs.LG
  17. Contextual Bandits with Continuous Actions: Smoothing, Zooming, and Adapting
    2019/02/05 by Akshay Krishnamurthy, John Langford, Krishnamurthy, Akshay +5 · 5 citations
    Computer Science · Decision Sciences · #Advanced Bandit Algorithms Research #Data Stream Mining Techniques #FOS: Computer and information sciences #Machine Learning (cs.LG) #Machine Learning (stat.ML) #Reinforcement Learning in Robotics
  18. Gone Fishing: Neural Active Learning with Fisher Embeddings
    2021/06/17 by Jordan T. Ash, Surbhi Goel, Ash, Jordan T. +5 · 5 citations
    Computer Science · #Adversarial Robustness in Machine Learning #FOS: Computer and information sciences #Machine Learning (cs.LG) #Machine Learning and Algorithms #Machine Learning and Data Classification
  19. Guaranteed Discovery of Control-Endogenous Latent States with Multi-Step Inverse Models
    2022/07/17 by Alex Lamb, Riashat Islam, Lamb, Alex +17 · 5 citations
    Computer Science · #Time Series Analysis and Forecasting #Reinforcement Learning in Robotics #AI-based Problem Solving and Planning
  20. Efficient First-Order Contextual Bandits: Prediction, Allocation, and Triangular Discrimination
    2021/07/05 by Dylan J. Foster, Foster, Dylan J., Akshay Krishnamurthy +1 · 4 citations
    Computer Science · Decision Sciences · #Advanced Bandit Algorithms Research #FOS: Computer and information sciences #FOS: Mathematics #Machine Learning (cs.LG) #Machine Learning (stat.ML) #Machine Learning and Algorithms #Machine Learning and Data Classification #Statistics Theory (math.ST)
  21. Streaming Active Learning with Deep Neural Networks
    2023/03/05 by Akanksha Saran, Saran, Akanksha, Safoora Yousefi +7 · 5 citations
    Computer Science · #Machine Learning and Algorithms #Data Stream Mining Techniques #Machine Learning and Data Classification
  22. Computational-Statistical Tradeoffs at the Next-Token Prediction Barrier: Autoregressive and Imitation Learning under Misspecification
    2025/02/18 by Dhruv Rohatgi, Rohatgi, Dhruv, Adam Block +7 · 1 voice · 7 citations
    Computer Science · #Natural Language Processing Techniques
  23. Optimism in Reinforcement Learning with Generalized Linear Function Approximation
    2019/12/09 by Yining Wang, Wang, Yining, Ruosong Wang +5 · 3 citations
    Computer Science · Decision Sciences · Engineering · #Advanced Bandit Algorithms Research #FOS: Computer and information sciences #Machine Learning (cs.LG) #Machine Learning (stat.ML) #Reinforcement Learning in Robotics #Smart Grid Energy Management
  24. Model-free Representation Learning and Exploration in Low-rank MDPs
    2021/02/14 by Aditya Modi, Jing‐Lin Chen, Modi, Aditya +7 · 6 citations
    Computer Science · Decision Sciences · Physics and Astronomy · #Reinforcement Learning in Robotics #Advanced Bandit Algorithms Research #Model Reduction and Neural Networks
  25. Efficient and Optimal Algorithms for Contextual Dueling Bandits under\n Realizability
    2021/11/24 by Aadirupa Saha, Akshay Krishnamurthy, Saha, Aadirupa +1 · 2 citations
    Computer Science · Decision Sciences · #Advanced Bandit Algorithms Research #Auction Theory and Applications #FOS: Computer and information sciences #Machine Learning (cs.LG) #Machine Learning and Algorithms #Optimization and Search Problems #Stochastic Gradient Optimization Techniques
  26. PAC Reinforcement Learning with Rich Observations
    2016/02/08 by Akshay Krishnamurthy, Krishnamurthy, Akshay, Alekh Agarwal +3 · 1 citation
    Computer Science · Decision Sciences · #Advanced Bandit Algorithms Research #Data Stream Mining Techniques #FOS: Computer and information sciences #Machine Learning (cs.LG) #Machine Learning (stat.ML) #Reinforcement Learning in Robotics
  27. On Oracle-Efficient PAC RL with Rich Observations
    2018/03/01 by Christoph Dann, Nan Jiang, Dann, Christoph +9 · 1 citation
    Computer Science · Decision Sciences · #Auction Theory and Applications #Evolutionary Algorithms and Applications #FOS: Computer and information sciences #Machine Learning (cs.LG) #Machine Learning (stat.ML) #Reinforcement Learning in Robotics
  28. Robust Dynamic Assortment Optimization in the Presence of Outlier Customers
    2019/10/09 by Xi Chen, Akshay Krishnamurthy, Chen, Xi +3 · 2 citations
    Decision Sciences · Computer Science · Business, Management and Accounting · #Advanced Bandit Algorithms Research #Optimization and Search Problems #Supply Chain and Inventory Management
  29. Butterfly Effects of SGD Noise: Error Amplification in Behavior Cloning and Autoregression
    2023/10/17 by Adam Block, Dylan J. Foster, Block, Adam +7 · 3 citations
    Computer Science · #FOS: Computer and information sciences #FOS: Mathematics #Machine Learning (cs.LG) #Machine Learning (stat.ML) #Optimization and Control (math.OC) #Reinforcement Learning in Robotics
  30. Kinematic State Abstraction and Provably Efficient Rich-Observation Reinforcement Learning
    2019/11/13 by Dipendra Misra, Mikael Henaff, Misra, Dipendra +5 · 1 citation
    Computer Science · Engineering · #Anomaly Detection Techniques and Applications #FOS: Computer and information sciences #Human Pose and Action Recognition #Machine Learning (cs.LG) #Machine Learning (stat.ML) #Robot Manipulation and Learning
  31. Semiparametric Contextual Bandits
    2018/03/12 by Akshay Krishnamurthy, Zhiwei Steven Wu, Krishnamurthy, Akshay +3 · 2 citations
    Decision Sciences · Computer Science · Engineering · #Advanced Bandit Algorithms Research #Reinforcement Learning in Robotics #Smart Grid Energy Management
  32. Provable RL with Exogenous Distractors via Multistep Inverse Dynamics
    2021/10/17 by Yonathan Efroni, Dipendra Misra, Efroni, Yonathan +7 · 1 citation
    Computer Science · Decision Sciences · #Advanced Bandit Algorithms Research #Data Stream Mining Techniques #FOS: Computer and information sciences #Machine Learning (cs.LG) #Reinforcement Learning in Robotics
  33. Oracle-Efficient Pessimism: Offline Policy Optimization in Contextual Bandits
    2023/06/13 by Lequn Wang, Wang, Lequn, Akshay Krishnamurthy +3 · 1 citation
    Computer Science · Decision Sciences · #Advanced Bandit Algorithms Research #FOS: Computer and information sciences #Machine Learning (cs.LG) #Machine Learning and Algorithms #Machine Learning and Data Classification
  34. Rich-Observation Reinforcement Learning with Continuous Latent Dynamics
    2024/05/29 by Yuda Song, Lili Wu, Song, Yuda +5 · 1 citation
    Computer Science · #Data Stream Mining Techniques #Domain Adaptation and Few-Shot Learning #FOS: Computer and information sciences #Machine Learning (cs.LG) #Machine Learning and Data Classification
  35. Wait, Wait, Wait... Why Do Reasoning Models Loop?
    2025/12/15 by Charilaos Pipis, Pipis, Charilaos, Garg, Shivam +8 · 1 citation
    Computer Science · Psychology · #Intelligent Tutoring Systems and Adaptive Learning #Topic Modeling #Visual and Cognitive Learning Processes