Jascha Sohl-Dickstein
- Levels of AGI for Operationalizing Progress on the Path to AGI
2023/11/04 by Meredith Ringel Morris, Jascha Sohl-Dickstein, Morris, Meredith Ringel +13 · 16 voices · 19 citations
#cs.AI
- Adversarial Examples that Fool both Computer Vision and Time-Limited Humans
2018/02/22 by Gamaleldin F. Elsayed, Shreya Shankar, Brian Cheung +4 · 4 voices · 6 citations
#cs.LG #cs.CV #q-bio.NC #stat.ML
- Adversarial Reprogramming of Neural Networks
2018/06/28 by Gamaleldin F. Elsayed, Ian Goodfellow, Jascha Sohl-Dickstein · 2 voices · 7 citations
#cs.LG #cs.CR #cs.CV #stat.ML
- Beyond the Imitation Game: Quantifying and extrapolating the capabilities of language models
2022/06/09 by Aarohi Srivastava, Abhinav Rastogi, Abhishek Rao +448 · 3 voices · 133 citations
#cs.CL #cs.AI #cs.CY #cs.LG #stat.ML
- Exponential expressivity in deep neural networks through transient chaos
2016/06/16 by Ben Poole, Subhaneil Lahiri, Poole, Ben +8 · 2 voices · 37 citations
Computer Science · Mathematics · Neuroscience · Physics and Astronomy · #Model Reduction and Neural Networks #Neural Networks and Applications #Neural dynamics and brain function #cond-mat.dis-nn #cs.LG #stat.ML
- Wide Neural Networks of Any Depth Evolve as Linear Models Under Gradient Descent
2019/02/18 by Jaehoon Lee, Lechao Xiao, Samuel S. Schoenholz +4 · 2 voices · 12 citations
Mathematics · Computer Science · #stat.ML #cs.LG
- The boundary of neural network trainability is fractal
2024/02/09 by Jascha Sohl-Dickstein, Jascha Sohl‐Dickstein, Sohl-Dickstein, Jascha · 3 voices · 6 citations
Computer Science · #Neural Networks and Applications
- Training LLMs over Neurally Compressed Text
2024/04/04 by Brian Lester, Jaehoon Lee, Lester, Brian +12 · 4 voices · 2 citations
Computer Science · #Handwritten Text Recognition Techniques #Mathematics, Computing, and Information Processing #Natural Language Processing Techniques #cs.CL #cs.LG
- Capacity and Trainability in Recurrent Neural Networks
2016/11/29 by Jasmine Collins, Jascha Sohl-Dickstein, Collins, Jasmine +4 · 1 voice · 3 citations
Computer Science · #Adversarial Robustness in Machine Learning #Neural Networks and Applications #Stochastic Gradient Optimization Techniques #cs.AI #cs.LG #cs.NE #stat.ML
- Scaling Exponents Across Parameterizations and Optimizers
2024/07/08 by Katie Everett, Lechao Xiao, Everett, Katie +19 · 2 voices · 20 citations
#cs.LG
- Fast large-scale optimization by unifying stochastic gradient and quasi-Newton methods
2013/11/08 by Jascha Sohl‐Dickstein, Jascha Sohl-Dickstein, Sohl-Dickstein, Jascha +4 · 1 voice · 3 citations
Computer Science · Engineering · #90C26 #Advanced Image Processing Techniques #FOS: Computer and information sciences #G.1.6 #Machine Learning (cs.LG) #Neural Networks and Applications #Sparse and Compressive Sensing Techniques #Stochastic Gradient Optimization Techniques #cs.LG
- VeLO: Training Versatile Learned Optimizers by Scaling Up
2022/11/17 by Luke Metz, J. Harrison, James Harrison +22 · 2 voices · 4 citations
Computer Science · #Advanced Neural Network Applications #Human Pose and Action Recognition #Machine Learning and Data Classification #cs.LG #math.OC #stat.ML
- Survey of Expressivity in Deep Neural Networks
2016/11/24 by Maithra Raghu, Ben Poole, Raghu, Maithra +7 · 1 voice
Computer Science · Mathematics · #FOS: Computer and information sciences #Machine Learning (cs.LG) #Machine Learning (stat.ML) #Neural and Evolutionary Computing (cs.NE) #cs.LG #cs.NE #stat.ML