Dynamical System Optimization

Emo Todorov

Abstract

We develop an optimization framework centered around a core idea: once a (parametric) policy is specified, control authority is transferred to the policy, resulting in an autonomous dynamical system. Thus we should be able to optimize policy parameters without further reference to controls or actions, and without directly using the machinery of approximate Dynamic Programming and Reinforcement Learning. Here we derive simpler algorithms at the autonomous system level, and show that they compute the same quantities as policy gradients and Hessians, natural gradients, proximal methods. Analogs to approximate policy iteration and off-policy learning are also available. Since policy parameters and other system parameters are treated uniformly, the same algorithms apply to behavioral cloning, mechanism design, system identification, learning of state estimators. Tuning of generative AI models is not only possible, but is conceptually closer to the present framework than to Reinforcement Learning.

Keywords

Artificial Intelligence & Data Science

📄 Full Paper Available as PDF

This paper is available as a downloadable PDF.

📄 Download PDF

Comments (0)

No comments yet. Be the first to comment.

Paper Details

Authors Emo Todorov
Published 2025-06-10
Category Artificial Intelligence And Data Science
Status Non-peer-reviewed Preprint
Language English
Word Count 142

Dynamical System Optimization

Abstract

Keywords

✨ AI Plain-English Summary

Comments (0)

Related Papers

An Efficient Algorithm for Computing Interventional Distributions in ...

Sparse matrix-variate Gaussian process blockmodels for network modeling

Hierarchical Maximum Margin Learning for Multi-Class Classification

Tightening MRF Relaxations with Planar Subproblems