Skip to content

Chulhee Yun

Publication record assembled from the DBLP archive of ranked conferences.

Papers indexed

26

Venues

4

Active years

2015–2025

Best venue rank

A*

Where they publish

Papers

26 indexed papers, newest first.

YearVenueTitleAuthors
2025ICLRArithmetic Transformers Can Length-Generalize in Both Operand Length and Count.Hanseul Cho, Jaeyoung Cha, Srinadh Bhojanapalli, Chulhee Yun
2025ICLRConvergence and Implicit Bias of Gradient Descent on Continual Linear Classification.Hyunji Jung, Hanseul Cho, Chulhee Yun
2025ICLRParameter Expanded Stochastic Gradient Markov Chain Monte Carlo.Hyunsu Kim, Giung Nam, Chulhee Yun, Hongseok Yang, Juho Lee
2025ICLRDoes SGD really happen in tiny subspaces?Minhak Song, Kwangjun Ahn, Chulhee Yun
2025ICMLLightweight Dataset Pruning without Full Training via Example Difficulty and Prediction Uncertainty.Yeseul Cho, Baekrok Shin, Changmin Kang, Chulhee Yun
2025ICMLProvable Benefit of Random Permutations over Uniform Sampling in Stochastic Coordinate Descent.Donghwa Kim, Jaewook Lee, Chulhee Yun
2025ICMLIncremental Gradient Descent with Small Epoch Counts is Surprisingly Slow on Ill-Conditioned Problems.Yujun Kim, Jaeyoung Cha, Chulhee Yun
2025ICMLUnderstanding Sharpness Dynamics in NN Training with a Minimalist Example: The Effects of Dataset Difficulty, Depth, Stochasticity, and More.Geonhui Yoo, Minhak Song, Chulhee Yun
2024ICLRLinear attention is (maybe) all you need (to understand Transformer optimization).Kwangjun Ahn, Xiang Cheng, Minhak Song, Chulhee Yun, Ali Jadbabaie, Suvrit Sra
2024ICMLFundamental Benefit of Alternating Updates in Minimax Optimization.Jaewook Lee, Hanseul Cho, Chulhee Yun
2023ICLRSGDA with shuffling: faster convergence for nonconvex-PŁ minimax optimization.Hanseul Cho, Chulhee Yun
2023ICMLTighter Lower Bounds for Shuffling SGD: Random Permutations and Beyond.Jaeyoung Cha, Jaewook Lee, Chulhee Yun
2023ICMLProvable Benefit of Mixup for Finding Optimal Decision Boundaries.Junsoo Oh, Chulhee Yun
2023ICMLOn the Training Instability of Shuffling SGD with Batch Normalization.David Xing Wu, Chulhee Yun, Suvrit Sra
2022ICLRMinibatch vs Local SGD with Shuffling: Tight Convergence Bounds and Beyond.Chulhee Yun, Shashank Rajput, Suvrit Sra
2021COLTProvable Memorization via Deep Neural Networks using Sub-linear Parameters.Sejun Park, Jaeho Lee, Chulhee Yun, Jinwoo Shin
2021COLTOpen Problem: Can Single-Shuffle SGD be Better than Reshuffling SGD and GD?Chulhee Yun, Suvrit Sra, Ali Jadbabaie
2021ICLRMinimum Width for Universal Approximation.Sejun Park, Chulhee Yun, Jaeho Lee, Jinwoo Shin
2021ICLRA unifying view on implicit bias in training linear neural networks.Chulhee Yun, Shankar Krishnan, Hossein Mobahi
2020ICLRAre Transformers universal approximators of sequence-to-sequence functions?Chulhee Yun, Srinadh Bhojanapalli, Ankit Singh Rawat, Sashank J. Reddi, Sanjiv Kumar
2020ICMLLow-Rank Bottleneck in Multi-head Attention Models.Srinadh Bhojanapalli, Chulhee Yun, Ankit Singh Rawat, Sashank J. Reddi, Sanjiv Kumar
2019ICLREfficiently testing local optimality and escaping saddles for ReLU networks.Chulhee Yun, Suvrit Sra, Ali Jadbabaie
2019ICLRSmall nonlinearities in activation functions create bad local minima in neural networks.Chulhee Yun, Suvrit Sra, Ali Jadbabaie
2018COLTMinimax Bounds on Stochastic Batched Convex Optimization.John C. Duchi, Feng Ruan, Chulhee Yun
2018ICLRGlobal Optimality Conditions for Deep Neural Networks.Chulhee Yun, Suvrit Sra, Ali Jadbabaie
2015ICASSPFace detection using Local Hybrid Patterns.Chulhee Yun, Donghoon Lee, Chang Dong Yoo