NSF PAR Search | NSF Public Access Repository

Note: When clicking on a Digital Object Identifier (DOI) number, you will be taken to an external site maintained by the publisher. Some full text articles may not yet be available without a charge during the embargo (administrative interval).
What is a DOI Number?

Some links on this page may take you to non-federal websites. Their policies may differ from this site.

Multi-Objective Covariance Matrix Adaptation MAP-Annealing

https://doi.org/10.1145/3712256.3726463

Zhao, Shihan; Nikolaidis, Stefanos (July 2025, ACM)

Free, publicly-accessible full text available July 13, 2026
Signal Temporal Logic-Guided Apprenticeship Learning

https://doi.org/10.1109/IROS58592.2024.10801924

Puranic, Aniruddh G; Deshmukh, Jyotirmoy V; Nikolaidis, Stefanos (October 2024, IEEE)

Apprenticeship learning crucially depends on effectively learning rewards, and hence control policies from user demonstrations. Of particular difficulty is the setting where the desired task consists of a number of sub-goals with temporal dependencies. The quality of inferred rewards and hence policies are typically limited by the quality of demonstrations, and poor inference of these can lead to undesirable outcomes. In this paper, we show how temporal logic specifications that describe high level task objectives, are encoded in a graph to define a temporal-based metric that reasons about behaviors of demonstrators and the learner agent to improve the quality of inferred rewards and policies. Through experiments on a diverse set of robot manipulator simulations, we show how our framework overcomes the drawbacks of prior literature by drastically improving the number of demonstrations required to learn a control policy.
more » « less
Full Text Available
Design of Communication Methods for Buoyancy Assisted Lightweight Legged Unit

Hsu, Ya-Chuan; Velentza, Anna-Maria; Nikolaidis, Stefanos (August 2024, Workshop on Trends in Socially Assistive Robotics (TSAR): Human-Centered Approach at the 33rd IEEE International Conference on Robot and Human Interactive Communication (RO-MAN).)

Full Text Available
Inferring Human Intent and Predicting Human Action in Human–Robot Collaboration

https://doi.org/10.1146/annurev-control-071223-105834

Hoffman, Guy; Bhattacharjee, Tapomayukh; Nikolaidis, Stefanos (July 2024, Annual Review of Control, Robotics, and Autonomous Systems)

Researchers in human–robot collaboration have extensively studied methods for inferring human intentions and predicting their actions, as this is an important precursor for robots to provide useful assistance. We review contemporary methods for intention inference and human activity prediction. Our survey finds that intentions and goals are often inferred via Bayesian posterior estimation and Markov decision processes that model internal human states as unobserved variables or represent both agents in a shared probabilistic framework. An alternative approach is to use neural networks and other supervised learning approaches to directly map observable outcomes to intentions and to make predictions about future human activity based on past observations. That said, due to the complexity of human intentions, existing work usually reasons about limited domains, makes unrealistic simplifications about intentions, and is mostly constrained to short-term predictions. This state of the art provides opportunity for future research that could include more nuanced models of intents, reason over longer horizons, and account for the human tendency to adapt.
more » « less
Full Text Available
Density Descent for Diversity Optimization

https://doi.org/10.1145/3638529.3654001

Lee, David H; Palaparthi, Anishalakshmi; Fontaine, Matthew C; Tjanaka, Bryon; Nikolaidis, Stefanos (July 2024, ACM)

Full Text Available
Enabling Adaptive Agent Training in Open-Ended Simulators by Targeting Diversity

Costales, Robby; Nikolaidis, Stefanos (March 2024, NeurIPS)

Full Text Available
Selecting Source Tasks for Transfer Learning of Human Preferences

https://doi.org/10.1109/LRA.2024.3415432

Nemlekar, Heramb; Sivagnanadasan, Naren; Banga, Subham; Dhanaraj, Neel; Gupta, Satyandra K; Nikolaidis, Stefanos (June 2024, IEEE Robotics and Automation Letters)

Full Text Available
Proximal Policy Gradient Arborescence for Quality Diversity Reinforcement Learning

Batra, Sumeet; Tjanaka, Bryon; Fontaine, Matthew; Petrenko, Aleksei; Nikolaidis, Stefanos; Sukhatme, Gaurav (May 2024, International Conference on Learning Representations (ICLR) 2024)

Full Text Available
Adapting Task Difficulty in a Cup-Stacking Rehabilitative Task

https://doi.org/10.1145/3610978.3640558

Daniilidis, Melina; Dennler, Nathaniel Steele; Matarić, Maja; Nikolaidis, Stefanos (March 2024, ACM)

Full Text Available
Covariance Matrix Adaptation MAP-Annealing: Theory and Experiments

https://doi.org/10.1145/3665336

Zhao, Shihan; Tjanaka, Bryon; Fontaine, Matthew_C; Nikolaidis, Stefanos (March 2025, ACM Transactions on Evolutionary Learning and Optimization)

Single-objective optimization algorithms search for the single highest quality solution with respect to an objective. Quality diversity (QD) optimization algorithms, such as Covariance Matrix Adaptation MAP-Elites (CMA-ME), search for a collection of solutions that are both high quality with respect to an objective and diverse with respect to specified measure functions. However, CMA-ME suffers from three major limitations highlighted by the QD community: prematurely abandoning the objective in favor of exploration, struggling to explore flat objectives, and having poor performance for low-resolution archives. We propose a new QD algorithm, CMA MAP-Annealing (CMA-MAE), and its differentiable QD variant, CMA-MAE via a Gradient Arborescence (CMA-MAEGA), that address all three limitations. We provide theoretical justifications for the new algorithm with respect to each limitation. Our theory informs our experiments, which support the theory and show that CMA-MAE achieves state-of-the-art performance and robustness on standard QD benchmark and reinforcement learning domains.
more » « less

« Prev Next »

Search for: All records