Total Intravenous Anesthesia (TIVA) automation requires rapid induction while avoiding excessive depth of anesthesia, a challenge exacerbated by strong interpatient pharmacodynamic variability. Fixed-gain PID controllers often fail to balance responsiveness and safety, whereas advanced methods such as MPC or neural networks introduce opacity that complicates regulatory certification and clinical acceptance. This work proposes a Real-Time Adaptive PID controller whose gains are scheduled through a deterministic policy derived via offline tabular Q-Learning. The learned gain-scheduling map links discretized tracking-error states to optimal PID parameters, enabling dynamic adaptation from induction to maintenance while preserving full interpretability and deterministic behavior, key requirements for European Medical Device Regulation approval. The approach was evaluated through Monte Carlo simulations on 100 virtual patients generated by randomizing Schnider and Hill model parameters. Compared to a conventional fixed-gain PID, the Adaptive RL-PID significantly reduced safety violations while maintaining deterministic, explainable behavior. The framework offers RL-driven performance improvements while remaining compatible with certification pathways for safety-critical medical controllers.

Explainable Adaptive PID Tuning with Reinforcement Learning in Anesthesia Induction / Poli, C., Menegatti, D., Giuseppi, A., Pietrabissa, A.. - (2026), pp. 581-587. (2026 34th Mediterranean Conference on Control and Automation (MED) Ancona; Italy ) [10.1109/med70602.2026.11598533].

Explainable Adaptive PID Tuning with Reinforcement Learning in Anesthesia Induction

Menegatti, Danilo;Giuseppi, Alessandro;Pietrabissa, Antonio
2026

Abstract

Total Intravenous Anesthesia (TIVA) automation requires rapid induction while avoiding excessive depth of anesthesia, a challenge exacerbated by strong interpatient pharmacodynamic variability. Fixed-gain PID controllers often fail to balance responsiveness and safety, whereas advanced methods such as MPC or neural networks introduce opacity that complicates regulatory certification and clinical acceptance. This work proposes a Real-Time Adaptive PID controller whose gains are scheduled through a deterministic policy derived via offline tabular Q-Learning. The learned gain-scheduling map links discretized tracking-error states to optimal PID parameters, enabling dynamic adaptation from induction to maintenance while preserving full interpretability and deterministic behavior, key requirements for European Medical Device Regulation approval. The approach was evaluated through Monte Carlo simulations on 100 virtual patients generated by randomizing Schnider and Hill model parameters. Compared to a conventional fixed-gain PID, the Adaptive RL-PID significantly reduced safety violations while maintaining deterministic, explainable behavior. The framework offers RL-driven performance improvements while remaining compatible with certification pathways for safety-critical medical controllers.
2026
2026 34th Mediterranean Conference on Control and Automation (MED)
Anesthesia Induction;Reinforcement Learning; PID
04 Pubblicazione in atti di convegno::04b Atto di convegno in volume
Explainable Adaptive PID Tuning with Reinforcement Learning in Anesthesia Induction / Poli, C., Menegatti, D., Giuseppi, A., Pietrabissa, A.. - (2026), pp. 581-587. (2026 34th Mediterranean Conference on Control and Automation (MED) Ancona; Italy ) [10.1109/med70602.2026.11598533].
File allegati a questo prodotto
File Dimensione Formato  
Menegatti_explainable-adaptive-PID_2026.pdf

solo gestori archivio

Tipologia: Versione editoriale (versione pubblicata con il layout dell'editore)
Licenza: Tutti i diritti riservati (All rights reserved)
Dimensione 1.28 MB
Formato Adobe PDF
1.28 MB Adobe PDF   Contatta l'autore

I documenti in IRIS sono protetti da copyright e tutti i diritti sono riservati, salvo diversa indicazione.

Utilizza questo identificativo per citare o creare un link a questo documento: https://hdl.handle.net/11573/1771680
Citazioni
  • ???jsp.display-item.citation.pmc??? ND
  • Scopus ND
  • ???jsp.display-item.citation.isi??? ND
social impact