Value Approximation for Two-Player General-Sum Differential Games with State Constraints

Zhang, Lei; Ghimire, Mukesh; Zhang, Wenlong; Xu, Zhe; Ren, Yi

Computer Science > Robotics

arXiv:2311.16520 (cs)

[Submitted on 28 Nov 2023 (v1), last revised 18 Apr 2024 (this version, v2)]

Title:Value Approximation for Two-Player General-Sum Differential Games with State Constraints

Authors:Lei Zhang, Mukesh Ghimire, Wenlong Zhang, Zhe Xu, Yi Ren

View PDF HTML (experimental)

Abstract:Solving Hamilton-Jacobi-Isaacs (HJI) PDEs numerically enables equilibrial feedback control in two-player differential games, yet faces the curse of dimensionality (CoD). While physics-informed neural networks (PINNs) have shown promise in alleviating CoD in solving PDEs, vanilla PINNs fall short in learning discontinuous solutions due to their sampling nature, leading to poor safety performance of the resulting policies when values are discontinuous due to state or temporal logic constraints. In this study, we explore three potential solutions to this challenge: (1) a hybrid learning method that is guided by both supervisory equilibria and the HJI PDE, (2) a value-hardening method where a sequence of HJIs are solved with increasing Lipschitz constant on the constraint violation penalty, and (3) the epigraphical technique that lifts the value to a higher dimensional state space where it becomes continuous. Evaluations through 5D and 9D vehicle and 13D drone simulations reveal that the hybrid method outperforms others in terms of generalization and safety performance by taking advantage of both the supervisory equilibrium values and costates, and the low cost of PINN loss gradients.

Comments:	Accepted to TRO
Subjects:	Robotics (cs.RO); Computer Science and Game Theory (cs.GT); Machine Learning (cs.LG)
Cite as:	arXiv:2311.16520 [cs.RO]
	(or arXiv:2311.16520v2 [cs.RO] for this version)
	https://doi.org/10.48550/arXiv.2311.16520

Submission history

From: Lei Zhang [view email]
[v1] Tue, 28 Nov 2023 04:58:41 UTC (9,916 KB)
[v2] Thu, 18 Apr 2024 04:53:16 UTC (10,752 KB)

Computer Science > Robotics

Title:Value Approximation for Two-Player General-Sum Differential Games with State Constraints

Submission history

Access Paper:

References & Citations

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators

Computer Science > Robotics

Title:Value Approximation for Two-Player General-Sum Differential Games with State Constraints

Submission history

Access Paper:

References & Citations

BibTeX formatted citation

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators