Aller au contenu principal
Profil bibliographique

Chia-Yu Hung

Informations fournies par OpenAlex. Research Africa ne déduit ni nationalité, ni poste, ni coordonnées personnelles.

7Publications signalées
54Citations signalées
1Affiliations récentes

Les institutions déclarées

Les domaines associés

Multimodal Machine Learning ApplicationsSpeech and Audio ProcessingLanguage, Metaphor, and CognitionMusic and Audio ProcessingMusic Technology and Sound Studies

Les publications récentes

Accès ouvert 2026 conference-paper OpenAlex

10 Open Challenges Steering the Future of Vision-Language-Action Models

Soujanya Poria, Navonil Majumder, Chia-Yu Hung, A. Bagherzadeh et autres

Due to their ability of follow natural language instructions, vision-language-action (VLA) models are increasingly preva- lent in the embodied AI arena, following the widespread suc- cess of their precursors—LLMs and VLMs. In this paper, we discuss 10 principal milestones in the ongoing …

sg, us (code pays fourni par la source)

1 citation Proceedings of the AAAI Conference on Artificial Intelligence
Accès ouvert 2025 preprint OpenAlex

10 Open Challenges Steering the Future of Vision-Language-Action Models

Soujanya Poria, Navonil Majumder, Chia-Yu Hung, A. Bagherzadeh et autres

Due to their ability of follow natural language instructions, vision-language-action (VLA) models are increasingly prevalent in the embodied AI arena, following the widespread success of their precursors -- LLMs and VLMs. In this paper, we discuss 10 principal milestones in the ongoing …

0 citations arXiv (Cornell University)
Accès ouvert 2025 preprint OpenAlex

NORA: A Small Open-Sourced Generalist Vision Language Action Model for Embodied Tasks

Chia-Yu Hung, Qi Sun, Pengfei Hong, Amir Zadeh et autres

Existing Visual-Language-Action (VLA) models have shown promising performance in zero-shot scenarios, demonstrating impressive task execution and reasoning capabilities. However, a significant challenge arises from the limitations of visual encoding, which can result in failures during tasks such as object grasping. Moreover, these …

0 citations arXiv (Cornell University)
Accès ouvert 2024 preprint OpenAlex

TangoFlux: Super Fast and Faithful Text to Audio Generation with Flow Matching and Clap-Ranked Preference Optimization

Chia-Yu Hung, Navonil Majumder, Zhifeng Kong, Ambuj Mehrish et autres

We introduce TangoFlux, an efficient Text-to-Audio (TTA) generative model with 515M parameters, capable of generating up to 30 seconds of 44.1kHz audio in just 3.7 seconds on a single A40 GPU. A key challenge in aligning TTA models lies in the difficulty …

0 citations arXiv (Cornell University)
2024 conference-paper OpenAlex

Contrastive Disentanglement for Authorship Attribution

Zhiqiang Hu, Thao Thanh Nguyen, Yujia Hu, Chia-Yu Hung et autres

Authorship Attribution (AA) seeks to determine the authorship of texts by examining distinctive writing styles. Although current AA methods have shown promising results, they often underperform in scenarios with significant topic shifts. This limitation arises from their inability to effectively separate topical …

sg (code pays fourni par la source)

2 citations
2023 conference-paper OpenAlex

Interpretable Sock Puppet Attribution

Chun Wei Seah, Chia-Yu Hung, Yujia Hu, Hai Leong Chieu et autres

The intentional spread of misinformation can have serious consequences in our society. This motivates us to address the task of identifying sock puppet accounts (i.e. fabricated online personas) created by individuals or organizations with the intention of deceiving their target audience. By …

sg (code pays fourni par la source)

2 citations

BNTIC News n’est pas le producteur de ces données. Les publications sont interrogées à la demande dans Crossref, OpenAIRE, DOAJ, Europe PMC, HAL, DataCite, AfricArXiv, ROR et la Banque mondiale, sans clé d’accès. OpenAlex reste optionnel. Aucun service payant n’est nécessaire et aucune donnée externe n’est enregistrée en base. Consulter les sources et leurs limites.