Koopman-Based Reinforcement Learning for LQ Control Gains Estimation of Quadrotors
| dc.contributor.author | Martini, Simone | |
| dc.contributor.author | Sonmez, Serhat | |
| dc.contributor.author | Stefanovic, Margareta | |
| dc.contributor.author | Rutherford, Matthew J. | |
| dc.contributor.author | Valavanis, Kimon P. | |
| dc.date.accessioned | 2025-11-16T19:34:12Z | |
| dc.date.issued | 2025 | |
| dc.department | İstanbul Medeniyet Üniversitesi | |
| dc.description | 2025 International Conference on Unmanned Aircraft Systems-ICUAS-Annual -- MAY 14-17, 2025 -- Charlotte, NC | |
| dc.description.abstract | In this research, Koopman operator theory is employed to achieve faster training time and improved performance of a reinforcement learning (RL) based linear quadratic controller (LQ). The proposed methodology, called K-RLLQ, is implemented for the trajectory tracking problem of a quadrotor UAV. Using the evolution of analytically derived Koopman generalized eigenfunctions allows for the embedding of quadrotor nonlinear dynamics into a quasi-linear model. Specifically, the resulting Koopman based quadrotor dynamics has linear state matrix and state dependent control matrix. Additionally, the obtained formulation is fully actuated, hence, compared to traditional model based hierarchical control the advantages are twofold: i) the controller can be formulated using linear control strategies in Koopman formulation which will result in a nonlinear control law in the original state space; ii) the trajectory tracking task can be achieved through a single control loop. Using this formulation, an RL agent is trained to estimate the controller parameters of a linear quadratic control law. Notably, it is shown that, using a reward function and observation space based on Koopman generalized eigenfunctions over the state space, leads to a considerably faster training time and improved overall performances. | |
| dc.description.sponsorship | Ministry of National Education of the Republic of Turkey on behalf of the Istanbul Medeniyet University, Turkey | |
| dc.description.sponsorship | Serhat Sonmez has been partially supported by the Ministry of National Education of the Republic of Turkey on behalf of the Istanbul Medeniyet University, Turkey. | |
| dc.description.sponsorship | Institute of Electrical and Electronics Engineers Inc | |
| dc.identifier.doi | 10.1109/ICUAS65942.2025.11007845 | |
| dc.identifier.endpage | 472 | |
| dc.identifier.isbn | 979-8-3315-1329-0 | |
| dc.identifier.isbn | 979-8-3315-1328-3 | |
| dc.identifier.issn | 2373-6720 | |
| dc.identifier.scopus | 2-s2.0-105007598006 | |
| dc.identifier.scopusquality | N/A | |
| dc.identifier.startpage | 465 | |
| dc.identifier.uri | https://doi.org/10.1109/ICUAS65942.2025.11007845 | |
| dc.identifier.uri | https://hdl.handle.net/20.500.14730/15272 | |
| dc.identifier.wos | WOS:001548686600062 | |
| dc.identifier.wosquality | N/A | |
| dc.indekslendigikaynak | Web of Science | |
| dc.indekslendigikaynak | Scopus | |
| dc.language.iso | en | |
| dc.publisher | Ieee | |
| dc.relation.ispartof | 2025 International Conference on Unmanned Aircraft Systems, Icuas | |
| dc.relation.publicationcategory | Konferans Öğesi - Uluslararası - Kurum Öğretim Elemanı | |
| dc.rights | info:eu-repo/semantics/closedAccess | |
| dc.snmz | KA_WOS_20250302 | |
| dc.title | Koopman-Based Reinforcement Learning for LQ Control Gains Estimation of Quadrotors | |
| dc.type | Conference Object |










