Surgan Jandial
- Thinking Fair and Slow: On the Efficacy of Structured Prompts for Debiasing Language Models
2024/05/16 by Shaz Furniturewala, Furniturewala, Shaz, Surgan Jandial +11 · 19 citations
Computer Science · #Natural Language Processing Techniques
- Introducing v0.5 of the AI Safety Benchmark from MLCommons
2024/04/18 by Bertie Vidgen, Adarsh Agrawal, Vidgen, Bertie +202 · 2 voices · 10 citations
Computer Science · #Adversarial Robustness in Machine Learning
- AdvGAN++ : Harnessing latent layers for adversary generation
2019/08/02 by Puneet Mangla, Mangla, Puneet, Surgan Jandial +5 · 1 citation
Computer Science · #Advanced Malware Detection Techniques #Adversarial Robustness in Machine Learning #Anomaly Detection Techniques and Applications #Computer Vision and Pattern Recognition (cs.CV) #FOS: Computer and information sciences #Machine Learning (cs.LG)
- All Should Be Equal in the Eyes of Language Models: Counterfactually Aware Fair Text Generation
2023/11/09 by Pragyan Banerjee, Abhinav Java, Banerjee, Pragyan +11 · 1 citation
Social Sciences · #Computation and Language (cs.CL) #Computers and Society (cs.CY) #Ethics and Social Impacts of AI #FOS: Computer and information sciences #Machine Learning (cs.LG)