AbstractMany large-scale machine learning problems -- clustering, non-parametric learning, kernel machines, etc. -- require selecting a small yet representative subset from a large dataset. Such problems can often be reduced to maximizing a submodular set function subject to various constraints. Classical approaches to
→