|
scholar.google.com
|
Google scholar page
|
|
arxiv.org
|
Smoothing DiLoCo with Primal Averaging for Faster Training of LLMs
|
|
arxiv.org
|
Stochastic Approximation with Block Coordinate Optimal Stepsizes
|
|
arxiv.org
|
Quantization through Piecewise-Affine Regularization: Optimization and Statistical Guarantees
|
|
openreview.net
|
Exploration from a Primal-Dual Lens: Value-Incentivized Actor-Critic Methods for Sample-Efficient Online RL
|
|
openreview.net
|
ParetoQ: Improving Scaling Laws in Extremely Low-bit LLM Quantization
|
|
openreview.net
|
LoRe: Personalizing LLMs via Low-Rank Reward Modeling
|
|
proceedings.mlr.press
|
PARQ: Piecewise-Affine Regularized Quantization
|
|
proceedings.mlr.press
|
Incentivize without Bonus: Provably Efficient Model-based Online Multi-agent RL for Markov Games
|
|
projecteuclid.org
|
Noisy recovery from random linear observations: Sharp minimax rates under elliptical constraints
|
|
arxiv.org
|
Dual Approximation Policy Optimization
|
|
arxiv.org
|
An Adaptive Stochastic Gradient Method with Non-negative Gauss-Newton Stepsizes
|
|
openreview.net
|
Linear Convergence of Natural Policy Gradient Methods with Log-Linear Policies
|
|
openreview.net
|
Faster Last-iterate Convergence of Policy Optimization in Zero-Sum Markov Games
|
|
pubsonline.informs.org
|
Stochastic Optimization with Decision-Dependent Distributions
|
|
link.springer.com
|
Stochastic variance-reduced prox-linear algorithms for nonconvex composite optimization
|
|
jmlr.org
|
On the convergence rates of policy gradient methods
|
|
proceedings.neurips.cc
|
BiT: Robustly Binarized Multi-distilled Transformer
|
|
proceedings.neurips.cc
|
Advances in Neural Information Processing Systems (NeurIPS) 35/, 2022.
|
|
proceedings.mlr.press
|
Federated Learning with Partial Model Personalization
|
|
openreview.net
|
FedShuffle: Recipes for Better Use of Local Work in Federated Learning
|
|
link.springer.com
|
On self-concordant barriers for generalized power cones
|
|
jmlr.org
|
From low probability to high confidence in stochastic convex optimization
|
|
link.springer.com
|
Accelerated Bregman proximal gradient methods for relatively smooth convex optimization
|
|
arxiv.org
|
MultiLevel Composite Stochastic Optimization via Nested Variance Reduction
|
|
arxiv.org
|
Adaptive stochastic variance reduction for subsampled Newton method with cubic regularization
|
|
arxiv.org
|
Statistical Adaptive Stochastic Gradient Methods
|
|
proceedings.mlr.press
|
Statistically Preconditioned Accelerated Gradient Method for Distributed Optimization
|
|
jmlr.org
|
DSCOVR: Randomized Primal-Dual Block Coordinate Algorithms for Asynchronous Distributed Optimization
|
|
papers.nips.cc
|
Using Statistics to Automate Stochastic Optimization
|
|
proceedings.neurips.cc
|
Understanding the Role of Momentum in Stochastic Gradient Methods
|
|
proceedings.neurips.cc
|
Stochastic Composite Gradient Method with Incremental Variance Reduction
|
|
proceedings.mlr.press
|
A Composite Randomized Incremental Gradient Method
|
|
papers.nips.cc
|
Coupled Variational Bayes via Optimization Embedding
|
|
papers.nips.cc
|
Learning SMaLL Predictors
|
|
proceedings.mlr.press
|
SBEED: Convergent Reinforcement Learning with Nonlinear Function Approximation
|
|
auai.org
|
Sparse Multi-Prototype Classification
|
|
arxiv.org
|
Communication-Efficient Distributed Optimization of Self-Concordant Empirical Loss
|
|
jmlr.org
|
Stochastic Primal-Dual Coordinate Method for Regularized Empirical Risk Minimization
|
|
arxiv.org
|
Variational Gram functions: convex analysis and optimization
|
|
arxiv.org
|
A Randomized Nonmonotone Block Proximal Gradient Method for a Class of Structured Nonlinear Programming
|
|
papers.nips.cc
|
Q-LDA: Uncovering Latent Patterns in Text-based Sequential Decision Processes
|
|
proceedings.mlr.press
|
Stochastic Variance Reduction Methods for Policy Evaluation
|
|
proceedings.mlr.press
|
Exploiting Strong Convexity from Data with Primal-Dual First-Order Algorithms
|
|
epubs.siam.org
|
An Accelerated Randomized Proximal Coordinate Gradient Method and its Application to Regularized Empirical Risk Minimization
|
|
arxiv.org
|
On the Complexity Analysis of Randomized Block-Coordinate Descent Methods
|
|
link.springer.com
|
An adaptive accelerated proximal gradient method and its homotopy continuation for sparse optimization
|
|
papers.nips.cc
|
End-to-end Learning of LDA by Mirror-Descent Back Propagation over a Deep Architecture
|
|
dl.acm.org
|
Scaling Up Stochastic Dual Coordinate Ascent
|
|
proceedings.mlr.press
|
Stochastic Primal-Dual Coordinate Method for Regularized Empirical Risk Minimization
|
|
proceedings.mlr.press
|
DiSCO: Distributed Optimization for Self-Concordant Empirical Loss
|
|
epubs.siam.org
|
A Proximal Stochastic Gradient Method with Progressive Variance Reduction
|
|
papers.nips.cc
|
An Accelerated Proximal Coordinate Gradient Method
|
|
aaai.org
|
Online Classification Using a Voted RDA Method
|
|
proceedings.mlr.press
|
An Adaptive Accelerated Proximal Gradient Method and its Homotopy Continuation for Sparse Optimization
|
|
epubs.siam.org
|
A Proximal-Gradient Homotopy Method for the Sparse Least-Squares Problem
|
|
jmlr.org
|
Optimal Distributed Online Prediction Using Mini-Batches
|
|
icml.cc
|
A Proximal-Gradient Homotopy Method for the L1-Regularized Least-Squares Problem
|
|
icml.cc
|
Hierarchical Classification via Orthogonal Transfer
|
|
icml.cc
|
Optimal Distributed Online Prediction
|
|
dl.acm.org
|
Distributed algorithms via gradient descent for fisher markets
|
|
jmlr.org
|
Dual Averaging Methods for Regularized Stochastic Learning and Online Optimization
|
|
ieeexplore.ieee.org
|
A Geometric Perspective of Large-Margin Training of Gaussian Models
|
|
link.springer.com
|
Learning to classify with missing and corrupted features
|
|
learningtheory.org
|
Optimal Algorithms for Online Convex Optimization with Multi-Point Bandit Feedback
|
|
epubs.siam.org
|
Fastest Mixing Markov Chain on Graphs with Symmetries
|
|
papers.nips.cc
|
Dual Averaging Method for Regularized Stochastic Learning and Online Optimization
|
|
usenix.org
|
Energy-Aware Server Provisioning and Load Dispatching for Connection-Intensive Internet Services
|
|
sciencedirect.com
|
Distributed average consensus with least-mean-square deviation
|
|
epubs.siam.org
|
The Fastest Mixing Markov Process on a Graph and a Connection to a Maximum Variance Unfolding Problem
|
|
ieeexplore.ieee.org
|
Cross-layer optimization of wireless networks using nonlinear column generation
|
|
link.springer.com
|
Optimal Scaling of a Gradient Method for Distributed Resource Allocation
|
|
tandfonline.com
|
Fastest Mixing Markov Chain on a Path
|
|
dl.acm.org
|
A duality view of spectral methods for dimensionality reduction
|
|
ieeexplore.ieee.org
|
A space-time diffusion scheme for peer-to-peer least-squares estimation
|
|
epubs.siam.org
|
Least-Squares Covariance Matrix Adjustment
|
|
projecteuclid.org
|
Symmetry Analysis of Reversible Markov Chains
|
|
link.springer.com
|
Joint Optimization of Wireless Communication and Networked Control Systems
|
|
dl.acm.org
|
A scheme for robust distributed sensor fusion based on average consensus
|
|
epubs.siam.org
|
Fastest Mixing Markov Chain on a Graph
|
|
sciencedirect.com
|
Fast linear iterations for distributed averaging
|
|
ieeexplore.ieee.org
|
Simultaneous routing and resource allocation via dual decomposition
|
|
ieeexplore.ieee.org
|
A decomposition approach to distributed analysis of networked systems
|
|
ieeexplore.ieee.org
|
Scheduling, routing and power allocation for fairness in wireless networks
|
|
ieeexplore.ieee.org
|
Joint optimization of communication rates and linear systems
|
|
ieeexplore.ieee.org
|
Fast linear iterations for distributed averaging
|
|
web.stanford.edu
|
Simultaneous Routing and Resource Allocation in CDMA Wireless Data Networks
|
|
ieeexplore.ieee.org
|
Control with random communication delays via a discrete-time jump system approach
|
|
github.com
|
jemdoc+MathJax
|