| 2016 | Active Uncertainty Calibration in Bayesian ODE Solvers. | Hans Kersting, Philipp Hennig |
| 2016 | Causal Inference by Minimizing the Dual Norm of Bias: Kernel Matching & Weighting Estimators for Causal Effects. | Nathan Kallus |
| 2016 | Hierarchical learning of grids of microtopics. | Nebojsa Jojic, Alessandro Perina, Dongwoo Kim |
| 2016 | Dynamical Kinds and their Discovery. | Benjamin C. Jantzen |
| 2016 | Scalable Nonparametric Bayesian Multilevel Clustering. | Viet Huynh, Dinh Q. Phung, Svetha Venkatesh, XuanLong Nguyen, Matthew D. Hoffman, Hung Hai Bui |
| 2016 | Efficient Feature Group Sequencing for Anytime Linear Prediction. | Hanzhang Hu, Alexander Grubb, J. Andrew Bagnell, Martial Hebert |
| 2016 | Structured Prediction: From Gaussian Perturbations to Linear-Time Principled Algorithms. | Jean Honorio, Tommi S. Jaakkola |
| 2016 | Bridging Heterogeneous Domains With Parallel Transport For Vision and Multimedia Applications. | Raghuraman Gopalan |
| 2016 | Model-Free Reinforcement Learning with Skew-Symmetric Bilinear Utilities. | Hugo Gilbert, Bruno Zanuttini, Paul Weng, Paolo Viappiani, Esther Nicart |
| 2016 | Training Neural Nets to Aggregate Crowdsourced Responses. | Alex Gaunt, Diana Borsa, Yoram Bachrach |
| 2016 | Degrees of Freedom in Deep Neural Networks. | Tianxiang Gao, Vladimir Jojic |
| 2016 | Scalable Joint Modeling of Longitudinal and Point Process Data for Disease Trajectory Prediction and Improving Management of Chronic Kidney Disease. | Joseph Futoma, Mark P. Sendak, Blake Cameron, Katherine A. Heller |
| 2016 | Quasi-Newton Hamiltonian Monte Carlo. | Tianfan Fu, Luo Luo, Zhihua Zhang |
| 2016 | Taming the Noise in Reinforcement Learning via Soft Updates. | Roy Fox, Ari Pakman, Naftali Tishby |
| 2016 | On the Theory and Practice of Privacy-Preserving Bayesian Data Analysis. | James R. Foulds, Joseph Geumlek, Max Welling, Kamalika Chaudhuri |
| 2016 | Bayesian Learning of Kernel Embeddings. | Seth R. Flaxman, Dino Sejdinovic, John P. Cunningham, Sarah Filippi |
| 2016 | Elliptical Slice Sampling with Expectation Propagation. | Francois Fagan, Jalaj Bhandari, John P. Cunningham |
| 2016 | Learning Network of Multivariate Hawkes Processes: A Time Series Approach. | Jalal Etesami, Negar Kiyavash, Kun Zhang, Kushagra Singhal |
| 2016 | Pruning Rules for Learning Parsimonious Context Trees. | Ralf Eggeling, Mikko Koivisto |
| 2016 | Online Bayesian Multiple Kernel Bipartite Ranking. | Changying Du, Changde Du, Guoping Long, Qing He, Yucheng Li |
| 2016 | Improving Predictive Accuracy Using Smart-Data rather than Big-Data: A Case Study of Soccer Teams' Evolving Performance. | Anthony C. Constantinou, Norman E. Fenton |
| 2016 | Optimal Stochastic Strongly Convex Optimization with a Logarithmic Number of Projections. | Jianhui Chen, Tianbao Yang, Qihang Lin, Lijun Zhang, Yi Chang |
| 2016 | Adversarial Inverse Optimal Control for General Imitation Learning Losses and Embodiment Transfer. | Xiangli Chen, Mathew Monfort, Brian D. Ziebart, Peter Carr |
| 2016 | Accelerated Stochastic Block Coordinate Gradient Descent for Sparsity Constrained Nonconvex Optimization. | Jinghui Chen, Quanquan Gu |
| 2016 | A Generative Block-Diagonal Model for Clustering. | Junxiang Chen, Jennifer G. Dy |