Tanabe, Takumi
- Stepwise Alignment for Constrained Language Model Policy Optimization
2024/04/17 by Wachi, Akifumi, Tran, Thien Q., Sato, Rei +2 · 5 citations
#Artificial Intelligence (cs.AI) #Computation and Language (cs.CL) #FOS: Computer and information sciences #Machine Learning (cs.LG)
- Max-Min Off-Policy Actor-Critic Method Focusing on Worst-Case Robustness to Model Misspecification
2022/11/07 by Takumi Tanabe, Rei Sato, Tanabe, Takumi +7 · 2 citations
Computer Science · Engineering · #Artificial Intelligence (cs.AI) #FOS: Computer and information sciences #Fuel Cells and Related Materials #I.2.6 #Machine Learning (cs.LG) #Reinforcement Learning in Robotics #Software Engineering Research
- Vulnerability Mitigation for Safety-Aligned Language Models via Debiasing
2025/02/04 by Thien Q. Tran, Tran, Thien Q., Akifumi Wachi +7 · 1 citation
Computer Science · Engineering · #Artificial Intelligence (cs.AI) #Computation and Language (cs.CL) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Safety Systems Engineering in Autonomy #Software Reliability and Analysis Research #Software Testing and Debugging Techniques