| 2017 | Finite-sum Composition Optimization via Variance Reduced Gradient Descent. | Xiangru Lian, Mengdi Wang, Ji Liu |
| 2017 | Binary and Multi-Bit Coding for Stable Random Projections. | Ping Li |
| 2017 | Less than a Single Pass: Stochastically Controlled Stochastic Gradient. | Lihua Lei, Michael I. Jordan |
| 2017 | Hierarchically-partitioned Gaussian Process Approximation. | Byung-Jun Lee, Jongmin Lee, Kee-Eung Kim |
| 2017 | Inference Compilation and Universal Probabilistic Programming. | Tuan Anh Le, Atilim Gunes Baydin, Frank D. Wood |
| 2017 | ASAGA: Asynchronous Parallel SAGA. | Rmi Leblond, Fabian Pedregosa, Simon Lacoste-Julien |
| 2017 | The End of Optimism? An Asymptotic Analysis of Finite-Armed Linear Bandits. | Tor Lattimore, Csaba Szepesvri |
| 2017 | Learning Cost-Effective and Interpretable Treatment Regimes. | Himabindu Lakkaraju, Cynthia Rudin |
| 2017 | Online Learning and Blackwell Approachability with Partial Monitoring: Optimal Convergence Rates. | Joon Kwon, Vianney Perchet |
| 2017 | Beta calibration: a well-founded and easily implemented improvement on logistic calibration for binary classifiers. | Meelis Kull, Telmo de Menezes e Silva Filho, Peter A. Flach |
| 2017 | Lipschitz Density-Ratios, Structured Data, and Data-driven Tuning. | Samory Kpotufe |
| 2017 | A Learning Theory of Ranking Aggregation. | Anna Korba, Stphan Clmenon, Eric Sibony |
| 2017 | Fast Bayesian Optimization of Machine Learning Hyperparameters on Large Datasets. | Aaron Klein, Stefan Falkner, Simon Bartels, Philipp Hennig, Frank Hutter |
| 2017 | Information Projection and Approximate Inference for Structured Sparse Variables. | Rajiv Khanna, Joydeep Ghosh, Russell A. Poldrack, Oluwasanmi Koyejo |
| 2017 | Scalable Greedy Feature Selection via Weak Submodularity. | Rajiv Khanna, Ethan R. Elenberg, Alexandros G. Dimakis, Sahand N. Negahban, Joydeep Ghosh |
| 2017 | Conjugate-Computation Variational Inference: Converting Variational Inference in Non-Conjugate Models to Inferences in Conjugate Models. | Mohammad Emtiyaz Khan, Wu Lin |
| 2017 | Stochastic Rank-1 Bandits. | Sumeet Katariya, Branislav Kveton, Csaba Szepesvri, Claire Vernade, Zheng Wen |
| 2017 | A Framework for Optimal Matching for Causal Inference. | Nathan Kallus |
| 2017 | Sequential Graph Matching with Sequential Monte Carlo. | Seong-Hwan Jun, Samuel W. K. Wong, James V. Zidek, Alexandre Bouchard-Ct |
| 2017 | Improved Strongly Adaptive Online Learning using Coin Betting. | Kwang-Sung Jun, Francesco Orabona, Stephen J. Wright, Rebecca Willett |
| 2017 | Robust and Efficient Computation of Eigenvectors in a Generalized Spectral Method for Constrained Clustering. | Chengming Jiang, Huiqing Xie, Zhaojun Bai |
| 2017 | Combinatorial Topic Models using Small-Variance Asymptotics. | Ke Jiang, Suvrit Sra, Brian Kulis |
| 2017 | Modal-set estimation with an application to clustering. | Heinrich Jiang, Samory Kpotufe |
| 2017 | Dynamic Collaborative Filtering With Compound Poisson Factorization. | Ghassen Jerfel, Mehmet Emin Basbug, Barbara E. Engelhardt |
| 2017 | Large-Scale Data-Dependent Kernel Approximation. | Catalin Ionescu, Alin-Ionut Popa, Cristian Sminchisescu |