Accelerating Stochastic Composition Optimization Mengdi Wang ⇤, Ji Liu
–Neural Information Processing Systems
Consider the stochastic composition optimization problem where the objective is a composition of two expected-value functions. We propose a new stochastic firstorder method, namely the accelerated stochastic compositional proximal gradient (ASC-PG) method, which updates based on queries to the sampling oracle using two different timescales. The ASC-PG is the first proximal gradient method for the stochastic composition problem that can deal with nonsmooth regularization penalty. We show that the ASC-PG exhibits faster convergence than the best known algorithms, and that it achieves the optimal sample-error complexity in several important special cases. We further demonstrate the application of ASC-PG to reinforcement learning and conduct numerical experiments.
Neural Information Processing Systems
Mar-12-2024, 14:43:05 GMT
- Country:
- North America > United States
- Pennsylvania (0.04)
- Europe
- Spain > Catalonia
- Barcelona Province > Barcelona (0.04)
- Russia > Central Federal District
- Moscow Oblast > Moscow (0.04)
- Netherlands > North Holland
- Amsterdam (0.04)
- Spain > Catalonia
- North America > United States
- Technology: