Goto

Collaborating Authors

 Industry








Accelerating Stochastic Composition Optimization

Neural Information Processing Systems

The popular stochastic gradient methods are well suited for minimizing expected-value objective functions or the sum of a large number of loss functions. Stochastic gradient methods find wide applications in estimation, online learning, and training of deep neural networks.




Structured Sparse Regression via Greedy Hard Thresholding

Neural Information Processing Systems

In this paper, we show that such NP-hard projections can not only be avoided by appealing to submodular optimization, but such methods come with strong theoretical guarantees even in the presence of poorly conditioned data (i.e. say when two features have