NSF PAR Search | NSF Public Access Repository

Note: When clicking on a Digital Object Identifier (DOI) number, you will be taken to an external site maintained by the publisher. Some full text articles may not yet be available without a charge during the embargo (administrative interval).
What is a DOI Number?

Some links on this page may take you to non-federal websites. Their policies may differ from this site.

Approximating Discontinuous Nash Equilibrial Values of Two-Player General-Sum Differential Games

https://doi.org/10.1109/ICRA48891.2023.10160219

Zhang, Lei; Ghimire, Mukesh; Zhang, Wenlong; Xu, Zhe; Ren, Yi (May 2023, 2023 IEEE International Conference on Robotics and Automation (ICRA))

Finding Nash equilibrial policies for two-player differential games requires solving Hamilton-Jacobi-Isaacs (HJI) PDEs. Self-supervised learning has been used to approximate solutions of such PDEs while circumventing the curse of dimensionality. However, this method fails to learn discontinuous PDE solutions due to its sampling nature, leading to poor safety performance of the resulting controllers in robotics applications when player rewards are discontinuous. This paper investigates two potential solutions to this problem: a hybrid method that leverages both supervised Nash equilibria and the HJI PDE, and a value-hardening method where a sequence of HJIs are solved with a gradually hardening reward. We compare these solutions using the resulting generalization and safety performance in two vehicle interaction simulation studies with 5D and 9D state spaces, respectively. Results show that with informative supervision (e.g., collision and near-collision demonstrations) and the low cost of self-supervised learning, the hybrid method achieves better safety performance than the supervised, self-supervised, and value hardening approaches on equal computational budget. Value hardening fails to generalize in the higher-dimensional case without informative supervision. Lastly, we show that the neural activation function needs to be continuously differentiable for learning PDEs and its choice can be case dependent.
more » « less
Full Text Available
When Shall I Estimate Your Intent? Costs and Benefits of Intent Inference in Multi-Agent Interactions

https://doi.org/10.23919/ACC53348.2022.9867155

Amatya, Sunny; Ghimire, Mukesh; Ren, Yi; Xu, Zhe; Zhang, Wenlong (June 2022, 2022 American Control Conference (ACC))

This paper addresses incomplete-information dynamic games, where reward parameters of agents are private. Previous studies have shown that online belief update is necessary for deriving equilibrial policies of such games, especially for high-risk games such as vehicle interactions. However, updating beliefs in real time is computationally expensive as it requires continuous computation of Nash equilibria of the sub-games starting from the current states. In this paper, we consider the triggering mechanism of belief update as a policy defined on the agents’ physical and belief states, and propose learning this policy through reinforcement learning (RL). Using a two-vehicle uncontrolled intersection case, we show that intermittent belief update via RL is sufficient for safe interactions, reducing the computation cost of updates by 59% when agents have full observations of physical states. Simulation results also show that the belief update frequency will increase as noise becomes more significant in measurements of the vehicle positions.
more » « less
Full Text Available
Targeted Attack on Deep RL-based Autonomous Driving with Learned Visual Patterns

https://doi.org/10.1109/ICRA46639.2022.9811574

Buddareddygari, Prasanth; Zhang, Travis; Yang, Yezhou; Ren, Yi (May 2022, 2022 International Conference on Robotics and Automation (ICRA))

Recent studies demonstrated the vulnerability of control policies learned through deep reinforcement learning against adversarial attacks, raising concerns about the application of such models to risk-sensitive tasks such as autonomous driving. Threat models for these demonstrations are limited to (1) targeted attacks through real-time manipulation of the agent's observation, and (2) untargeted attacks through manipulation of the physical environment. The former assumes full access to the agent's states/observations at all times, while the latter has no control over attack outcomes. This paper investigates the feasibility of targeted attacks through visually learned patterns placed on physical objects in the environment, a threat model that combines the practicality and effectiveness of the existing ones. Through analysis, we demonstrate that a pre-trained policy can be hijacked within a time window, e.g., performing an unintended self-parking, when an adversarial object is present. To enable the attack, we adopt an assumption that the dynamics of both the environment and the agent can be learned by the attacker. Lastly, we empirically show the effectiveness of the proposed attack on different driving scenarios, perform a location robustness test, and study the tradeoff between the attack strength and its effectiveness Code is available at https://github.com/ASU-APG/ Targeted-Physical-Adversarial-Attacks-on-AD
more » « less
Full Text Available
Injecting Semantic Concepts Into End-to-End Image Captioning

Fang, Zhiyuan; Wang, Jianfeng; Hu, Xiaowei; Liang, Lin; Gan, Zhe; Wang, Lijuan; Yang, Yezhou and (January 2022, Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR))

Full Text Available
When Shall I Be Empathetic? The Utility of Empathetic Parameter Estimation in Multi-Agent Interactions

https://doi.org/10.1109/ICRA48506.2021.9561079

Chen, Yi; Zhang, Lei; Merry, Tanner; Amatya, Sunny; Zhang, Wenlong; Ren, Yi (May 2021, 2021 IEEE International Conference on Robotics and Automation (ICRA))

Full Text Available
Low to High Dimensional Modality Hallucination Using Aggregated Fields of View

https://doi.org/10.1109/LRA.2020.2970679

Gunasekar, Kausic; Qiu, Qiang; Yang, Yezhou (April 2020, IEEE Robotics and Automation Letters)

Full Text Available
Enabling Courteous Vehicle Interactions through Game-based and Dynamics-aware Intent Inference

https://doi.org/10.1109/TIV.2019.2955897

Wang, Yiwei; Ren, Yi; Elliott, Steven; Zhang, Wenlong (January 2020, IEEE Transactions on Intelligent Vehicles)

Full Text Available

Search for: All records