Automating inspection and listing processes
We develop vision language action (VLA) models that incorporate new modalities such as tactile sensors, as well as methods that prevent camera viewpoints from affecting model learning. Through these technologies, we aim to replace human-performed tasks such as unpacking and inspecting items with robotic systems and automate the listing process.
