vix.ing · top · new · best · stats · spec

Alfano, Carlo

  1. A Novel Framework for Policy Mirror Descent with General Parameterization and Linear Convergence
    2023/01/30 by Carlo Alfano, Alfano, Carlo, Rui Yuan +3 · 4 citations
    Computer Science · Engineering · #Advanced Memory and Neural Computing #FOS: Computer and information sciences #FOS: Mathematics #Fuel Cells and Related Materials #Machine Learning (cs.LG) #Machine Learning (stat.ML) #Optimization and Control (math.OC) #Reinforcement Learning in Robotics #Statistics Theory (math.ST)
  2. Linear Convergence for Natural Policy Gradient with Log-linear Policy Parametrization
    2022/09/30 by Carlo Alfano, Alfano, Carlo, Patrick Rebeschini +1 · 2 citations
    Computer Science · #Stochastic Gradient Optimization Techniques #Reinforcement Learning in Robotics #Age of Information Optimization
  3. Learning mirror maps in policy mirror descent
    2024/02/07 by Carlo Alfano, Sebastian Towers, Alfano, Carlo +7 · 1 voice · 1 citation
    Business, Management and Accounting · Computer Science · #FOS: Computer and information sciences #FOS: Mathematics #Machine Learning (cs.LG) #Machine Learning (stat.ML) #Machine Learning and Data Classification #Optimization and Control (math.OC) #Political Influence and Corporate Strategies #Topic Modeling
  4. Meta-Learning Objectives for Preference Optimization
    2024/11/10 by Carlo Alfano, Alfano, Carlo, Silvia Sapora +7 · 2 citations
    Decision Sciences · #Artificial Intelligence (cs.AI) #Decision-Making and Behavioral Economics #FOS: Computer and information sciences #Machine Learning (cs.LG) #Machine Learning (stat.ML)