Chi, Jianfeng
- The Llama 3 Herd of Models
2024/07/31 by Grattafiori, Aaron, Dubey, Abhimanyu, Jauhri, Abhinav +556 · 2794 citations
#Artificial Intelligence (cs.AI) #Computation and Language (cs.CL) #Computer Vision and Pattern Recognition (cs.CV) #FOS: Computer and information sciences
- Llama Guard: LLM-based Input-Output Safeguard for Human-AI Conversations
2023/12/07 by Hakan Inan, Inan, Hakan, Kartikeya Upasani +19 · 174 citations
Computer Science · Health Professions · #Hate Speech and Cyberbullying Detection #Interpreting and Communication in Healthcare
- Persistent Pre-Training Poisoning of LLMs
2024/10/17 by Yiming Zhang, Javier Rando, Zhang, Yiming +13 · 3 voices · 21 citations
#cs.CR #cs.AI
- Llama Guard 3 Vision: Safeguarding Human-AI Image Understanding Conversations
2024/11/15 by Jianfeng Chi, Ujjwal Karn, Chi, Jianfeng +17 · 37 citations
Earth and Planetary Sciences · #3D Surveying and Cultural Heritage #Computation and Language (cs.CL) #Computer Vision and Pattern Recognition (cs.CV) #FOS: Computer and information sciences
- Llama Guard 3-1B-INT4: Compact and Efficient Safeguard for Human-AI Conversations
2024/11/18 by Igor Fedorov, Kate Plawiak, Fedorov, Igor +37 · 2 voices · 5 citations
#cs.DC #cs.AI
- Backtracking Improves Generation Safety
2024/09/22 by Zhang, Yiming, Chi, Jianfeng, Nguyen, Hailey +4 · 10 citations
#Artificial Intelligence (cs.AI) #Computation and Language (cs.CL) #FOS: Computer and information sciences #Machine Learning (cs.LG)
- PolicyQA: A Reading Comprehension Dataset for Privacy Policies
2020/10/06 by Wasi Uddin Ahmad, Jianfeng Chi, Ahmad, Wasi Uddin +5 · 3 citations
Computer Science · Psychology · Social Sciences · #Access Control and Trust #Computation and Language (cs.CL) #FOS: Computer and information sciences #Mental Health via Writing #Topic Modeling
- BadMerging: Backdoor Attacks Against Model Merging
2024/08/14 by Jinghuai Zhang, Jianfeng Chi, Zhang, Jinghuai +9 · 6 citations
Computer Science · Social Sciences · #Access Control and Trust #Cryptography and Security (cs.CR) #FOS: Computer and information sciences #Formal Methods in Verification #Machine Learning (cs.LG) #Security and Verification in Computing
- FFB: A Fair Fairness Benchmark for In-Processing Group Fairness Methods
2023/06/15 by Han, Xiaotian, Chi, Jianfeng, Chen, Yu +4 · 4 citations
#Artificial Intelligence (cs.AI) #Computers and Society (cs.CY) #FOS: Computer and information sciences #Machine Learning (cs.LG)
- Towards Understanding the Fragility of Multilingual LLMs against Fine-Tuning Attacks
2024/10/23 by Poppi, Samuele, Yong, Zheng-Xin, He, Yifei +4 · 4 citations
#Artificial Intelligence (cs.AI) #Computation and Language (cs.CL) #Cryptography and Security (cs.CR) #FOS: Computer and information sciences #Machine Learning (cs.LG)
- Intent Classification and Slot Filling for Privacy Policies
2021/01/01 by Ahmad, Wasi Uddin, Chi, Jianfeng, Le, Tu +3 · 1 citation
#Computation and Language (cs.CL) #FOS: Computer and information sciences
- Understanding and Mitigating Accuracy Disparity in Regression
2021/02/24 by Chi, Jianfeng, Tian, Yuan, Gordon, Geoffrey J. +1 · 1 citation
#Computers and Society (cs.CY) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Machine Learning (stat.ML)
- Towards Return Parity in Markov Decision Processes
2021/11/19 by Chi, Jianfeng, Shen, Jian, Dai, Xinyi +3 · 1 citation
#Artificial Intelligence (cs.AI) #Computers and Society (cs.CY) #FOS: Computer and information sciences #Machine Learning (cs.LG) #Machine Learning (stat.ML)
- Hybrid Batch Attacks: Finding Black-box Adversarial Examples with Limited Queries
2019/08/19 by Suya, Fnu, Chi, Jianfeng, Evans, David +1 · 1 citation
#Cryptography and Security (cs.CR) #FOS: Computer and information sciences
- EAVE: Efficient Product Attribute Value Extraction via Lightweight Sparse-layer Interaction
2024/06/10 by Yang Li, Qifan Wang, Yang, Li +17 · 1 citation
Engineering · #Computation and Language (cs.CL) #FOS: Computer and information sciences #Industrial Vision Systems and Defect Detection