<?xml version="1.1" encoding="utf-8"?>
<article xsi:noNamespaceSchemaLocation="http://jats.nlm.nih.gov/publishing/1.1/xsd/JATS-journalpublishing1-mathml3.xsd" dtd-version="1.1" xmlns:xlink="http://www.w3.org/1999/xlink" xmlns:mml="http://www.w3.org/1998/Math/MathML" xmlns:xsi="http://www.w3.org/2001/XMLSchema-instance"><front><journal-meta><journal-id journal-id-type="publisher-id">JERA</journal-id><journal-title-group><journal-title>Journal of Electronic Research and Application</journal-title></journal-title-group><issn>2208-3502</issn><eissn>2208-3510</eissn><publisher><publisher-name>Bio-Byword Scientific Publishing Pty. Ltd.</publisher-name></publisher></journal-meta><article-meta><article-id pub-id-type="doi">10.26689/jera.v10i4.14898</article-id><article-categories><subj-group subj-group-type="heading"><subject>Article</subject></subj-group></article-categories><title>3D Object Detection Algorithm based on Improved PV-RCNN</title><url>https://artdesignp.com/journal/JERA/10/4/10.26689/jera.v10i4.14898</url><author>ZhaoZhicheng,JiaShijie</author><pub-date pub-type="publication-year"><year>2026</year></pub-date><volume>10</volume><issue>4</issue><history><date date-type="pub"><published-time>2026-05-21</published-time></date></history><abstract>In autonomous driving perception, point cloud-based 3D object detection plays an important role. This task still faces two challenges in long-range and small-object detection: loss of fine details and weak context modeling. To solve these problems, this paper proposes HFA-RCNN based on PV-RCNN. The method adds an encoder-decoder structure to the 3D sparse convolution backbone. This design improves multi-scale context modeling and preserves more detailed features. In the BEV feature generation stage, the method also designs a spatial-frequency aggregation network. This network combines complementary information from the spatial domain and the frequency domain. This design improves feature representation. Results on the KITTI dataset show that the proposed method preserves strong detection performance for the Car category and further improves detection accuracy for the Pedestrian and Cyclist categories. These results confirm the effectiveness of the method in long-range and small-object detection.</abstract><keywords/></article-meta></front><body/><back><ref-list><ref id="B1" content-type="article"><label>1</label><element-citation publication-type="journal"><p>Zhou Y, Tuzel O, 2018, VoxelNet: End-to-End Learning for Point Cloud based 3D Object Detection, Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition, 4490–4499.</p><pub-id pub-id-type="doi"/></element-citation></ref><ref id="B2" content-type="article"><label>2</label><element-citation publication-type="journal"><p>Yan Y, Mao Y, Li B, 2018, Second: Sparsely Embedded Convolutional Detection. Sensors, 18(10): 3337.</p><pub-id pub-id-type="doi"/></element-citation></ref><ref id="B3" content-type="article"><label>3</label><element-citation publication-type="journal"><p>Qi C, Su H, Mo K, et al., 2017, Pointnet: Deep Learning on Point Sets for 3D Classification and Segmentation, Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition, 652–660.</p><pub-id pub-id-type="doi"/></element-citation></ref><ref id="B4" content-type="article"><label>4</label><element-citation publication-type="journal"><p>Shi S, Wang X, Li H, 2019, PointRCNN: 3D Object Proposal Generation and Detection from Point Cloud, Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition, 770–779.</p><pub-id pub-id-type="doi"/></element-citation></ref><ref id="B5" content-type="article"><label>5</label><element-citation publication-type="journal"><p>Shi S, Guo C, Jiang L, et al., 2020, PV-RCNN: Point-Voxel Feature Set Abstraction for 3D Object Detection, Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition, 10526–10535.</p><pub-id pub-id-type="doi"/></element-citation></ref><ref id="B6" content-type="article"><label>6</label><element-citation publication-type="journal"><p>Geiger A, Lenz P, Urtasun R, 2012, Are We Ready for Autonomous Driving? The KITTI Vision Benchmark Suite, Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition, 3354–3361.</p><pub-id pub-id-type="doi"/></element-citation></ref><ref id="B7" content-type="article"><label>7</label><element-citation publication-type="journal"><p>Ku J, Mozifian M, Lee J, et al., 2018, Joint 3D Proposal Generation and Object Detection from View Aggregation, 2018 IEEE/RSJ International Conference on Intelligent Robots and Systems (IROS), 1–8.</p><pub-id pub-id-type="doi"/></element-citation></ref><ref id="B8" content-type="article"><label>8</label><element-citation publication-type="journal"><p>Lang A, Vora S, Caesar H, et al., 2019, PointPillars: Fast Encoders for Object Detection from Point Clouds, Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition, 12689–12697.</p><pub-id pub-id-type="doi"/></element-citation></ref><ref id="B9" content-type="article"><label>9</label><element-citation publication-type="journal"><p>Shi S, Wang Z, Shi J, et al., 2020, Part-A^2 Net: 3D Part-Aware and Aggregation Network for Object Detection from Point Cloud, Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR), 13338–13347.</p><pub-id pub-id-type="doi"/></element-citation></ref><ref id="B10" content-type="article"><label>10</label><element-citation publication-type="journal"><p>Pan X, Xia Z, Song S, et al., 2021, 3D Object Detection with Pointformer, Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR), 7463–7472.</p><pub-id pub-id-type="doi"/></element-citation></ref></ref-list></back></article>
