Aller au contenu principal
2024 article

LVP: Leverage Virtual Points in Multimodal Early Fusion for 3-D Object Detection

10Citations signalées, ce qui n’est pas une note de qualité
3Institutions déclarées
2Pays d’affiliation déclarés

Rattachement africain : cn, ca. Niveau de preuve : code pays fourni par la source.

Le résumé fourni par la source

Due to the sparsity and occlusion of point clouds, pure point cloud detection has limited effectiveness in detecting such samples. Researchers have been actively exploring the fusion of multimodal data, attempting to address the bottleneck issue based on LiDAR. In particular, virtual points, generated through depth completion from front-view RGB image, offer the potential for better integration with point clouds. Nevertheless, recent approaches fuse these two modalities in the region of interest (RoI), which limits the fusion effectiveness due to the inaccurate RoI region issue in the point cloud’s branch, especially in hard samples. To overcome it and unleash the potential of virtual points, while combining late fusion, we present leverage virtual point (LVP), a high-performance 3-D object detector which LVPs in early fusion to enhance the quality of RoI generation. LVP consists of three early fusion modules: virtual points painting (VPP), virtual points auxiliary (VPA), and virtual points completion (VPC) to achieve point-level fusion and global-level fusion. The integration of these modules effectively improves occlusion handling and improves the detection of distant small objects. In the KITTI benchmark, LVP achieves 85.45% 3-D mAP. As for large dataset nuScenes, we could improve the detection accuracy of large objects by compensating for errors in depth estimation. Without whistles and bells, these results establish LVP as an impressive solution for a 3-D outdoor object detection algorithm.

Ce résumé expose les affirmations des auteurs. BNTIC ne l’interprète pas comme une validation indépendante des résultats.

Le contrôle bibliographique ouvert

DOI retrouvé dans Crossref DOI retrouvé ; titre concordant.

Titre Crossref
LVP: Leverage Virtual Points in Multimodal Early Fusion for 3-D Object Detection
Date Crossref
01/01/2025
Éditeur
Institute of Electrical and Electronics Engineers (IEEE)
Type
journal-article

Ce recoupement confirme des métadonnées liées au DOI. Il ne confirme ni la méthode ni les conclusions de l’étude, et il ne compte pas comme une seconde source scientifique indépendante.

Où se fait cette recherche

  • Jimei University Office of Science and Research pays non établi dans la notice
    Université ou école supérieure
  • Beijing Jiaotong University pays non établi dans la notice
    Université ou école supérieure
  • University of Waterloo Department of Geography and Environmental Management and the Department of System Design Engineering pays non établi dans la notice
    Université ou école supérieure
  • Computer Engineering College pays non établi dans la notice
    Université ou école supérieure
  • School of Computer and Information Technology pays non établi dans la notice
    Université ou école supérieure

Office of Science and Research — Jimei University, Beijing Jiaotong University et Department of Geography and Environmental Management and the Department of System Design Engineering — University of Waterloo, avec 2 autres affiliations.

Une affiliation ne permet pas de déduire la nationalité d’un auteur.

Les sujets associés

Robotics and Sensor-Based LocalizationIndustrial Vision Systems and Defect DetectionAdvanced Neural Network Applications

BNTIC News n’est pas le producteur de ces données. Les publications sont interrogées à la demande dans Crossref, OpenAIRE, DOAJ, Europe PMC, HAL, DataCite, AfricArXiv, ROR et la Banque mondiale, sans clé d’accès. OpenAlex reste optionnel. Aucun service payant n’est nécessaire et aucune donnée externe n’est enregistrée en base. Consulter les sources et leurs limites.