FlexiFly: Interfacing the Physical World with Foundation Models Empowered by Reconfigurable Drone Systems

Zhao, Minghui; Xia, Junxi; Hou, Kaiyuan; Liu, Yanchen; Xia, Stephen; Jiang, Xiaofan

doi:10.1145/3715014.3722081

Citation Details

FlexiFly: Interfacing the Physical World with Foundation Models Empowered by Reconfigurable Drone Systems

Foundation models (FM) have shown immense human-like capabilities for generating digital media. However, foundation models that can freely sense, interact, and actuate the physical domain is far from being realized. This is due to 1) requiring dense deployments of sensors to fully cover and analyze large spaces, while 2) events often being localized to small areas, making it difficult for FMs to pinpoint relevant areas of interest relevant to the current task. We propose FlexiFly, a platform that enables FMs to “zoom in” and analyze relevant areas with higher granularity to better understand the physical environment and carry out tasks. FlexiFly accomplishes by introducing 1) a novel image segmentation technique that aids in identifying relevant locations and 2) a modular and reconfigurable sensing and actuation drone platform that FMs can actuate to “zoom in” with relevant sensors and actuators. We demonstrate through real smart home deployments that FlexiFly enables FMs and LLMs to complete diverse tasks up to 85% more successfully. FlexiFly is critical step towards FMs and LLMs that can naturally interface with the physical world. more »

Award ID(s):: 1943396

PAR ID:: 10592096

Author(s) / Creator(s):: Zhao, Minghui; Xia, Junxi; Hou, Kaiyuan; Liu, Yanchen; Xia, Stephen; Jiang, Xiaofan

Publisher / Repository:: ACM

Date Published:: 2025-05-06

ISBN:: 9798400714795

Page Range / eLocation ID:: 463 to 476

Format(s):: Medium: X

Location:: UC Irvine Student Center. Irvine CA USA

Sponsoring Org:: National Science Foundation

Free Publicly Accessible Full Text
Accepted Manuscript
Conference Paper:
https://doi.org/10.1145/3715014.3722081

More Like this