Accès ouvert
2026
preprint
OpenAlex
Dongyoung Kim, Huiwon Jang, M. Koo, Suhyeok Jang et autres
While Vision-Language-Action models (VLAs) have shown remarkable progress toward human-like generalist robotic policies through the versatile intelligence (i.e. broad scene understanding and language-conditioned generalization) inherited from pre-trained Vision-Language Models, they still struggle with complex real-world tasks requiring broader functional capabilities (e.g. motion …
Accès ouvert
2026
preprint
OpenAlex
Jimin Lee, Huiwon Jang, M. Koo, Jungwoo Park et autres
Humans understand and interact with the real world by relying on diverse physical feedback beyond visual perception. Motivated by this, recent approaches attempt to incorporate physical sensory signals into Vision-Language-Action models (VLAs). However, they typically focus on a single type of physical …
Accès ouvert
2025
preprint
OpenAlex
Tae‐Young Kim, Jimin Lee, M. Koo, Dongyoung Kim et autres
Vision-Language-Action (VLA) models have shown strong capabilities in robot manipulation by leveraging rich representations from pre-trained Vision-Language Models (VLMs). However, their representations arguably remain suboptimal, lacking sensitivity to robotic signals such as control actions and proprioceptive information. To address the issue, we …
Accès ouvert
2025
preprint
OpenAlex
M. Koo, Subin Kim, Sangkyung Kwak, Jae Hyun Nam et autres
Text-to-image diffusion models have significantly improved the seamless integration of visual text into diverse image contexts. Recent approaches further improve control over font styles through fine-tuning with predefined font dictionaries. However, adapting unseen fonts outside the preset is computationally expensive, often requiring …
2024
conference-abstract
OpenAlex
N. Solanki, Nicholas Wanner, Brittany Beck, Monica Labadia et autres
us
(code pays fourni par la source)
Accès ouvert
2023
article
OpenAlex
Megan M. Lowery, Nicholas S. Hill, Lu Wang, Erika B. Rosenzweig et autres
us
(code pays fourni par la source)
2023
conference-abstract
OpenAlex
S.A.A. Comhair, Kewal Asosingh, Yanhong Hou, Nicholas Wanner et autres
us, gb
(code pays fourni par la source)