Aller au contenu principal
Profil bibliographique

Luca Zappella

Informations fournies par OpenAlex. Research Africa ne déduit ni nationalité, ni poste, ni coordonnées personnelles.

57Publications signalées
896Citations signalées
2Affiliations récentes

Les institutions déclarées

Les domaines associés

Topic ModelingAdvanced Vision and ImagingDomain Adaptation and Few-Shot LearningNatural Language Processing TechniquesVideo Surveillance and Tracking Methods

Les publications récentes

Accès ouvert 2026 preprint OpenAlex

The Communication Bottleneck: A Round-Trip Study of Tree-Structured Expression Serialization in Language Models

Xavier Suau, Alex Ferrando de las Morenas, Luca Zappella, Samy Bengio

When language models reason in chain-of-thought or exchange free-text intermediates, they serialize structured information into natural language. How much tree-structured compositional content survives this bottleneck? We propose a round-trip protocol that answers this question empirically for tree-structured expressions. A generator converts a …

0 citations arXiv (Cornell University)
Accès ouvert 2026 preprint OpenAlex

On The Effectiveness-Fluency Trade-Off In LLM Conditioning: A Systematic Study

Iuri Macocco, Pau Rodríguez, Arno Blaas, Luca Zappella et autres

Controlling the output of Large Language Models (LLMs) is a central challenge for their reliable deployment, yet a clear understanding of the involved trade-offs remains elusive. Current approaches to conditioning are often evaluated with a narrow focus on their effectiveness at injecting …

0 citations arXiv (Cornell University)
Accès ouvert 2026 preprint OpenAlex

On The Effectiveness-Fluency Trade-Off In LLM Conditioning: A Systematic Study

Iuri Macocco, Pau Rodríguez, Arno Blaas, Luca Zappella et autres

Controlling the output of Large Language Models (LLMs) is a central challenge for their reliable deployment, yet a clear understanding of the involved trade-offs remains elusive. Current approaches to conditioning are often evaluated with a narrow focus on their effectiveness at injecting …

es, gb (code pays fourni par la source)

0 citations arXiv (Cornell University)
Accès ouvert 2026 preprint OpenAlex

HyperTransport: Amortized Conditioning of T2I Generative Models

Valentino Maiorca, Eleonora Gualdoni, Xavier Suau, Marco Cuturi et autres

As foundation models grow in capability, the ability to efficiently and reliably control their behavior becomes critical. Fine-tuning these models can be costly, and while prompting can be practical for controllability, it remains fragile due to models' high sensitivity to exact prompt …

0 citations arXiv (Cornell University)
Accès ouvert 2026 preprint OpenAlex

HyperTransport: Amortized Conditioning of T2I Generative Models

Valentino Maiorca, Eleonora Gualdoni, Xavier Suau, Marco Cuturi et autres

As foundation models grow in capability, the ability to efficiently and reliably control their behavior becomes critical. Fine-tuning these models can be costly, and while prompting can be practical for controllability, it remains fragile due to models' high sensitivity to exact prompt …

0 citations arXiv (Cornell University)
Accès ouvert 2026 preprint OpenAlex

Uncertainty Quantification for LLM Function-Calling

Zihuiwen Ye, Lukas Aichberger, Michael Kirchhof, Sinead Williamson et autres

Large Language Models (LLMs) are increasingly deployed to autonomously solve real-world tasks. A key ingredient for this is the LLM Function-Calling paradigm, a widely used approach for equipping LLMs with tool-use capabilities. However, an LLM calling functions incorrectly can have severe implications, …

0 citations arXiv (Cornell University)
Accès ouvert 2026 preprint OpenAlex

Uncertainty Quantification for LLM Function-Calling

Zihuiwen Ye, Lukas Aichberger, Michael Kirchhof, Sinead Williamson et autres

Large Language Models (LLMs) are increasingly deployed to autonomously solve real-world tasks. A key ingredient for this is the LLM Function-Calling paradigm, a widely used approach for equipping LLMs with tool-use capabilities. However, an LLM calling functions incorrectly can have severe implications, …

gb (code pays fourni par la source)

0 citations arXiv (Cornell University)
Accès ouvert 2026 preprint OpenAlex

Attention to Mamba: A Recipe for Cross-Architecture Distillation

Abhinav Moudgil, Ningyuan Teresa Huang, Eeshan Gunesh Dhekane, Pau Riera Rodriguez et autres

State Space Models (SSMs) such as Mamba have become a popular alternative to Transformer models, due to their reduced memory consumption and higher throughput at generation compared to their Attention-based counterparts. On the other hand, the community has built up a considerable …

0 citations arXiv (Cornell University)
Accès ouvert 2026 preprint OpenAlex

Attention to Mamba: A Recipe for Cross-Architecture Distillation

Abhinav Moudgil, Ningyuan Teresa Huang, Eeshan Gunesh Dhekane, Pau Riera Rodriguez et autres

State Space Models (SSMs) such as Mamba have become a popular alternative to Transformer models, due to their reduced memory consumption and higher throughput at generation compared to their Attention-based counterparts. On the other hand, the community has built up a considerable …

us, Algérie (code pays fourni par la source)

0 citations arXiv (Cornell University)
Accès ouvert 2026 preprint OpenAlex

The Design Space of Tri-Modal Masked Diffusion Models

Louis Bethune, Victor Turrisi, Bruno Mlodozeniec, Pau Rodriguez Lopez et autres

Discrete diffusion models have emerged as strong alternatives to autoregressive language models, with recent work initializing and fine-tuning a base unimodal model for bimodal generation. Diverging from previous approaches, we introduce the first tri-modal masked diffusion model pretrained from scratch on text, …

0 citations arXiv (Cornell University)
Accès ouvert 2026 preprint OpenAlex

The Design Space of Tri-Modal Masked Diffusion Models

Louis Bethune, Victor Turrisi, Bruno Mlodozeniec, Pau Rodriguez Lopez et autres

Discrete diffusion models have emerged as strong alternatives to autoregressive language models, with recent work initializing and fine-tuning a base unimodal model for bimodal generation. Diverging from previous approaches, we introduce the first tri-modal masked diffusion model pretrained from scratch on text, …

0 citations arXiv (Cornell University)
Accès ouvert 2026 preprint OpenAlex

GenCtrl -- A Formal Controllability Toolkit for Generative Models

Emily Cheng, Carmen Amo Alonso, Federico Danieli, Arno Blaas et autres

As generative models become ubiquitous, there is a critical need for fine-grained control over the generation process. Yet, while controlled generation methods from prompting to fine-tuning proliferate, a fundamental question remains unanswered: are these models truly controllable in the first place? In …

0 citations arXiv (Cornell University)

BNTIC News n’est pas le producteur de ces données. Les publications sont interrogées à la demande dans Crossref, OpenAIRE, DOAJ, Europe PMC, HAL, DataCite, AfricArXiv, ROR et la Banque mondiale, sans clé d’accès. OpenAlex reste optionnel. Aucun service payant n’est nécessaire et aucune donnée externe n’est enregistrée en base. Consulter les sources et leurs limites.