NSF PAR Search | NSF Public Access Repository

Note: When clicking on a Digital Object Identifier (DOI) number, you will be taken to an external site maintained by the publisher. Some full text articles may not yet be available without a charge during the embargo (administrative interval).
What is a DOI Number?

Some links on this page may take you to non-federal websites. Their policies may differ from this site.

ODYSSEE: Oyster Detection Yielded by Sensor Systems on Edge Electronics

Lin, Xiaomin; Mange, Vivek; Suresh, Arjun; Neuberger, Bernhard; Palnitkar, Aadi; Campbell, Brendan; Williams, Alan; Baxevani, Kleio; Mallette, Jeremy; Vera, Alhim; et al (May 2025, IEEE International Conference on Robotics and Automation (ICRA))

Oysters are a vital keystone species in coastal ecosystems, providing significant economic, environmental, and cultural benefits. As the importance of oysters grows, so does the relevance of autonomous systems for their detection and monitoring. However, current monitoring strategies often rely on destructive methods. While manual identification of oysters from video footage is non-destructive, it is time-consuming, requires expert input, and is further complicated by the challenges of the underwater environment. To address these challenges, we propose a novel pipeline using stable diffusion to augment a collected real dataset with photorealistic synthetic data. This method enhances the dataset used to train a YOLOv10-based vision model. The model is then deployed and tested on an edge platform; Aqua2, an Autonomous Underwater Vehicle (AUV), achieving a state-of-the-art 0.657 mAP@50 for oyster detection.
more » « less
Free, publicly-accessible full text available May 28, 2026
A Linear Time and Space Local Point Cloud Geometry Encoder via Vectorized Kernel Mixture (VecKM)

Yuan, Dehao; Fermuller, Cornelia; Rabbani, Tahseen; Huang, Furong; Aloimonos, Yiannis (January 2025, JMLR.org)
Salakhutdinov, Ruslan; Kolter, Zico; Heller, Katherine; Weller, Adrian; Nuria, Jonathan; Scarlett, Oliver; Berkenkamp, Felix (Ed.)
We propose VecKM, a local point cloud geometry encoder that is descriptive and efficient to compute. VecKM leverages a unique approach by vectorizing a kernel mixture to represent the local point cloud. Such representation's descriptiveness is supported by two theorems that validate its ability to reconstruct and preserve the similarity of the local shape. Unlike existing encoders down-sampling the local point cloud, VecKM constructs the local geometry encoding using all neighboring points, producing a more descriptive encoding. Moreover, VecKM is efficient to compute and scalable to large point cloud inputs: VecKM reduces the memory cost from (n2 + nKd) to (nd + np); and reduces the major runtime cost from computing nK MLPs to n MLPs, where n is the size of the point cloud, K is the neighborhood size, d is the encoding dimension, and p is a marginal factor. The efficiency is due to VecKM's unique factorizable property that eliminates the need of explicitly grouping points into neighbors. In the normal estimation task, VecKM demonstrates not only 100× faster inference speed but also highest accuracy and strongest robustness. In classification and segmentation tasks, integrating VecKM as a preprocessing module achieves consistently better performance than the PointNet, PointNet++, and point transformer baselines, and runs consistently faster by up to 10 times.
more » « less
Free, publicly-accessible full text available January 3, 2026
Active Human Pose Estimation via an Autonomous UAV Agent

https://doi.org/10.1109/IROS58592.2024.10801780

Chen, Jingxi; He, Botao; Singh, Chahat Deep; Fermüller, Cornelia; Aloimonos, Yiannis (October 2024, IEEE)

Full Text Available
Embodiment: Self-Supervised Depth Estimation Based on Camera Models

https://doi.org/10.1109/IROS58592.2024.10801343

Zhang, Jinchang; Reddy, Praveen Kumar; Wong, Xue-Iuan; Aloimonos, Yiannis; Lu, Guoyu (October 2024, IEEE)

Full Text Available
Minimal perception: enabling autonomy in resource-constrained robots

https://doi.org/10.3389/frobt.2024.1431826

Singh, Chahat Deep; He, Botao; Fermüller, Cornelia; Metzler, Christopher; Aloimonos, Yiannis (September 2024, Frontiers in Robotics and AI)

The rapidly increasing capabilities of autonomous mobile robots promise to make them ubiquitous in the coming decade. These robots will continue to enhance efficiency and safety in novel applications such as disaster management, environmental monitoring, bridge inspection, and agricultural inspection. To operate autonomously without constant human intervention, even in remote or hazardous areas, robots must sense, process, and interpret environmental data using only onboard sensing and computation. This capability is made possible by advancements in perception algorithms, allowing these robots to rely primarily on their perception capabilities for navigation tasks. However, tiny robot autonomy is hindered mainly by sensors, memory, and computing due to size, area, weight, and power constraints. The bottleneck in these robots lies in the real-time perception in resource-constrained robots. To enable autonomy in robots of sizes that are less than 100 mm in body length, we draw inspiration from tiny organisms such as insects and hummingbirds, known for their sophisticated perception, navigation, and survival abilities despite their minimal sensor and neural system. This work aims to provide insights into designing a compact and efficient minimal perception framework for tiny autonomous robots from higher cognitive to lower sensor levels.
more » « less
Full Text Available
Vector Symbolic Sub-objects Classifiers as Manifold Analogues

https://doi.org/10.1109/IJCNN60899.2024.10651219

Faraone, Renato; Sutor, Peter; Fermüller, Cornelia; Aloimonos, Yiannis (June 2024, IEEE)

Full Text Available
AcTExplore: Active Tactile Exploration on Unknown Objects

https://doi.org/10.1109/ICRA57147.2024.10611667

Shahidzadeh, Amir-Hossein; Yoo, Seong Jong; Mantripragada, Pavan; Singh, Chahat Deep; Fermüller, Cornelia; Aloimonos, Yiannis (May 2024, IEEE)

Full Text Available
Decodable and Sample Invariant Continuous Object Encoder

Yuan, Dehao; Huang, Furong; Fermuller, Cornelia; Aloimonos, Yiannis (January 2024, The Twelfth International Conference on Learning Representations)
Microsaccade-inspired event camera for robotics

https://doi.org/10.1126/scirobotics.adj8124

He, Botao; Wang, Ze; Zhou, Yuan; Chen, Jingxi; Singh, Chahat Deep; Li, Haojia; Gao, Yuman; Shen, Shaojie; Wang, Kaiwei; Cao, Yanjun; et al (May 2024, Science Robotics)

Neuromorphic vision sensors or event cameras have made the visual perception of extremely low reaction time possible, opening new avenues for high-dynamic robotics applications. These event cameras’ output is dependent on both motion and texture. However, the event camera fails to capture object edges that are parallel to the camera motion. This is a problem intrinsic to the sensor and therefore challenging to solve algorithmically. Human vision deals with perceptual fading using the active mechanism of small involuntary eye movements, the most prominent ones called microsaccades. By moving the eyes constantly and slightly during fixation, microsaccades can substantially maintain texture stability and persistence. Inspired by microsaccades, we designed an event-based perception system capable of simultaneously maintaining low reaction time and stable texture. In this design, a rotating wedge prism was mounted in front of the aperture of an event camera to redirect light and trigger events. The geometrical optics of the rotating wedge prism allows for algorithmic compensation of the additional rotational motion, resulting in a stable texture appearance and high informational output independent of external motion. The hardware device and software solution are integrated into a system, which we call artificial microsaccade–enhanced event camera (AMI-EV). Benchmark comparisons validated the superior data quality of AMI-EV recordings in scenarios where both standard cameras and event cameras fail to deliver. Various real-world experiments demonstrated the potential of the system to facilitate robotics perception both for low-level and high-level vision tasks.
more » « less
Full Text Available
Ajna: Generalized deep uncertainty for minimal perception on parsimonious robots

https://doi.org/10.1126/scirobotics.add5139

Sanket, Nitin J.; Singh, Chahat Deep; Fermüller, Cornelia; Aloimonos, Yiannis (August 2023, Science Robotics)

Robots are active agents that operate in dynamic scenarios with noisy sensors. Predictions based on these noisy sensor measurements often lead to errors and can be unreliable. To this end, roboticists have used fusion methods using multiple observations. Lately, neural networks have dominated the accuracy charts for perception-driven predictions for robotic decision-making and often lack uncertainty metrics associated with the predictions. Here, we present a mathematical formulation to obtain the heteroscedastic aleatoric uncertainty of any arbitrary distribution without prior knowledge about the data. The approach has no prior assumptions about the prediction labels and is agnostic to network architecture. Furthermore, our class of networks, Ajna, adds minimal computation and requires only a small change to the loss function while training neural networks to obtain uncertainty of predictions, enabling real-time operation even on resource-constrained robots. In addition, we study the informational cues present in the uncertainties of predicted values and their utility in the unification of common robotics problems. In particular, we present an approach to dodge dynamic obstacles, navigate through a cluttered scene, fly through unknown gaps, and segment an object pile, without computing depth but rather using the uncertainties of optical flow obtained from a monocular camera with onboard sensing and computation. We successfully evaluate and demonstrate the proposed Ajna network on four aforementioned common robotics and computer vision tasks and show comparable results to methods directly using depth. Our work demonstrates a generalized deep uncertainty method and demonstrates its utilization in robotics applications.
more » « less
Full Text Available

« Prev Next »

Search for: All records