Accès ouvert
2026
preprint
OpenAlex
Ting Huang, Zhenyu Zhang, W E I-Xing Huang, Jian Yang et autres
Video spatial reasoning is essential for navigation-oriented perception and long-video question answering, where models must infer spatial relations across long horizons under changing viewpoints. However, existing multimodal large language models (MLLMs) remain largely semantic-centric, and often fail to reliably aggregate consistent spatial …
Accès ouvert
2026
preprint
OpenAlex
Thomas Adam, Fengpeng An, Costas Andreopoulos, Giuseppe Andronico et autres
The Jiangmen Underground Neutrino Observatory (JUNO) collaboration has completed the construction of the 20,000-ton liquid scintillator detector and the associated muon veto detector system. To meet the physics objectives, the materials used in the detector must exhibit low radioactive contamination. The single-event …
fr, cn, it, de, tw, be, ca
(code pays fourni par la source)
Accès ouvert
2026
preprint
OpenAlex
Andy Dai, Zexue He, Zhenyu Zhang, Alex Pentland et autres
Current benchmarks for language models primarily evaluate execution on fully specified tasks. However, real user tasks are often ambiguous. Users arrive with incomplete, exploratory, or even inconsistent goals, requiring the assistant to first determine the intended task before carrying it out. We …
us
(code pays fourni par la source)