Accès ouvert
2026
conference-paper
OpenAlex
Soujanya Poria, Navonil Majumder, Chia-Yu Hung, A. Bagherzadeh et autres
Due to their ability of follow natural language instructions, vision-language-action (VLA) models are increasingly preva- lent in the embodied AI arena, following the widespread suc- cess of their precursors—LLMs and VLMs. In this paper, we discuss 10 principal milestones in the ongoing …
sg, us
(code pays fourni par la source)
Accès ouvert
2025
preprint
OpenAlex
Soujanya Poria, Navonil Majumder, Chia-Yu Hung, A. Bagherzadeh et autres
Due to their ability of follow natural language instructions, vision-language-action (VLA) models are increasingly prevalent in the embodied AI arena, following the widespread success of their precursors -- LLMs and VLMs. In this paper, we discuss 10 principal milestones in the ongoing …
Accès ouvert
2025
preprint
OpenAlex
Chia-Yu Hung, Qi Sun, Pengfei Hong, Amir Zadeh et autres
Existing Visual-Language-Action (VLA) models have shown promising performance in zero-shot scenarios, demonstrating impressive task execution and reasoning capabilities. However, a significant challenge arises from the limitations of visual encoding, which can result in failures during tasks such as object grasping. Moreover, these …
Accès ouvert
2024
preprint
OpenAlex
Chia-Yu Hung, Navonil Majumder, Zhifeng Kong, Ambuj Mehrish et autres
We introduce TangoFlux, an efficient Text-to-Audio (TTA) generative model with 515M parameters, capable of generating up to 30 seconds of 44.1kHz audio in just 3.7 seconds on a single A40 GPU. A key challenge in aligning TTA models lies in the difficulty …
2024
conference-paper
OpenAlex
Navonil Majumder, Chia-Yu Hung, Deepanway Ghosal, Wei-Ning Hsu et autres
Peer Reviewed
sg, us
(code pays fourni par la source)
2024
conference-paper
OpenAlex
Zhiqiang Hu, Thao Thanh Nguyen, Yujia Hu, Chia-Yu Hung et autres
Authorship Attribution (AA) seeks to determine the authorship of texts by examining distinctive writing styles. Although current AA methods have shown promising results, they often underperform in scenarios with significant topic shifts. This limitation arises from their inability to effectively separate topical …
sg
(code pays fourni par la source)
2023
conference-paper
OpenAlex
Chun Wei Seah, Chia-Yu Hung, Yujia Hu, Hai Leong Chieu et autres
The intentional spread of misinformation can have serious consequences in our society. This motivates us to address the task of identifying sock puppet accounts (i.e. fabricated online personas) created by individuals or organizations with the intention of deceiving their target audience. By …
sg
(code pays fourni par la source)