Makkuva, Ashok Vardhan
- Optimal transport mapping via input convex neural networks
2019/08/28 by Ashok Vardhan Makkuva, Amirhossein Taghvaei, Makkuva, Ashok Vardhan +5 · 25 citations
Computer Science · #Domain Adaptation and Few-Shot Learning #FOS: Computer and information sciences #Machine Learning (cs.LG) #Machine Learning (stat.ML) #Machine Learning and ELM #Stochastic Gradient Optimization Techniques
- Attention with Markov: A Framework for Principled Analysis of Transformers via Markov Chains
2024/02/06 by Makkuva, Ashok Vardhan, Bondaschi, Marco, Girish, Adway +4 · 10 citations
#Computation and Language (cs.CL) #FOS: Computer and information sciences #Information Theory (cs.IT) #Machine Learning (cs.LG) #Machine Learning (stat.ML)
- Transformers on Markov Data: Constant Depth Suffices
2024/07/25 by Nived Rajaraman, Rajaraman, Nived, Marco Bondaschi +7 · 12 citations
Computer Science · Mathematics · #Computation and Language (cs.CL) #FOS: Computer and information sciences #Information Theory (cs.IT) #Machine Learning (cs.LG) #Machine Learning (stat.ML) #Markov Chains and Monte Carlo Methods #Stochastic Gradient Optimization Techniques
- KO codes: Inventing Nonlinear Encoding and Decoding for Reliable Wireless Communication via Deep-learning
2021/08/29 by Makkuva, Ashok Vardhan, Liu, Xiyang, Jamali, Mohammad Vahid +3 · 5 citations
#Artificial Intelligence (cs.AI) #FOS: Computer and information sciences #Information Theory (cs.IT)
- Local to Global: Learning Dynamics and Effect of Initialization for Transformers
2024/06/05 by Ashok Vardhan Makkuva, Marco Bondaschi, Makkuva, Ashok Vardhan +11 · 7 citations
Psychology · #FOS: Computer and information sciences #Information Theory (cs.IT) #Innovative Teaching and Learning Methods #Machine Learning (cs.LG) #Machine Learning (stat.ML)
- CRISP: Curriculum based Sequential Neural Decoders for Polar Code Family
2022/10/01 by Hebbar, S Ashwin, Nadkarni, Viraj, Makkuva, Ashok Vardhan +3 · 3 citations
#Artificial Intelligence (cs.AI) #FOS: Computer and information sciences #Information Theory (cs.IT) #Machine Learning (cs.LG)
- Fundamental Limits of Prompt Compression: A Rate-Distortion Framework for Black-Box Language Models
2024/07/22 by Alliot Nagle, Nagle, Alliot, Adway Girish +9 · 3 citations
Computer Science · #Embedded Systems Design Techniques #Parallel Computing and Optimization Techniques
- Equivalence of additive-combinatorial linear inequalities for Shannon entropy and differential entropy
2016/01/27 by Makkuva, Ashok Vardhan, Wu, Yihong · 1 citation
#FOS: Computer and information sciences #Information Theory (cs.IT)
- From Markov to Laplace: How Mamba In-Context Learns Markov Chains
2025/02/14 by Bondaschi, Marco, Rajaraman, Nived, Wei, Xiuying +5 · 4 citations
#Artificial Intelligence (cs.AI) #FOS: Computer and information sciences #Information Theory (cs.IT) #Machine Learning (cs.LG)
- Reed-Muller Subcodes: Machine Learning-Aided Design of Efficient Soft Recursive Decoding
2021/02/02 by Jamali, Mohammad Vahid, Liu, Xiyang, Makkuva, Ashok Vardhan +3 · 1 citation
#FOS: Computer and information sciences #Information Theory (cs.IT)
- LASER: Linear Compression in Wireless Distributed Optimization
2023/10/19 by Ashok Vardhan Makkuva, Makkuva, Ashok Vardhan, Marco Bondaschi +8 · 1 citation
Computer Science · Decision Sciences · Engineering · #Advanced Bandit Algorithms Research #Artificial Intelligence (cs.AI) #FOS: Computer and information sciences #Information Theory (cs.IT) #Machine Learning (cs.LG) #Neural and Evolutionary Computing (cs.NE) #Sparse and Compressive Sensing Techniques #Stochastic Gradient Optimization Techniques