The KITTI Vision Benchmark Suite

	Method	Setting	Code	Moderate	Easy	Hard	Runtime	Environment
1	ViKIENet			98.06 %	98.63 %	93.21 %	0.04 s	1 core @ 2.5 Ghz (C/C++)

2	LVP			97.84 %	98.70 %	93.07 %	0.04 s	1 core @ 2.5 Ghz (C/C++)
Y. Chen, G. Cai, Z. Song, Z. Liu, B. Zeng, J. Li and Z. Wang: LVP: Leverage Virtual Points in Multi- modal Early Fusion for 3D Object Detection. IEEE Transactions on Geoscience and Remote Sensing 2024.
3	SCEMF			97.61 %	98.64 %	94.63 %	1 s	1 core @ 2.5 Ghz (C/C++)

4	UDeerPEP		code	97.57 %	98.42 %	95.08 %	0.1 s	1 core @ 2.5 Ghz (C/C++)
Z. Dong, H. Ji, X. Huang, W. Zhang, X. Zhan and J. Chen: PeP: a Point enhanced Painting method for unified point cloud tasks. 2023.
5	OGMMDet		code	97.54 %	98.44 %	92.88 %	0.01 s	1 core @ 2.5 Ghz (C/C++)

6	ViKIENet-R			97.40 %	95.89 %	92.63 %	0.06 s	1 core @ 2.5 Ghz (C/C++)

7	RM3D			97.35 %	95.93 %	92.67 %	0.18 s	1 core @ 2.5 Ghz (C/C++)

8	VirConv-S		code	97.27 %	98.00 %	94.53 %	0.09 s	1 core @ 2.5 Ghz (C/C++)
H. Wu, C. Wen, S. Shi and C. Wang: Virtual Sparse Convolution for Multimodal 3D Object Detection. CVPR 2023.
9	ANM		code	97.23 %	97.67 %	94.59 %	ANM	ANM

10	MuStD			97.21 %	97.91 %	94.04 %	67 ms	>8 cores @ 2.5 Ghz (Python)

11	GraR-VoI		code	96.38 %	96.81 %	91.20 %	0.07 s	1 core @ 2.5 Ghz (Python + C/C++)
H. Yang, Z. Liu, X. Wu, W. Wang, W. Qian, X. He and D. Cai: Graph R-CNN: Towards Accurate 3D Object Detection with Semantic-Decorated Local Graph. ECCV 2022.
12	VirConv-T		code	96.38 %	98.93 %	93.56 %	0.09 s	1 core @ 2.5 Ghz (C/C++)
H. Wu, C. Wen, S. Shi and C. Wang: Virtual Sparse Convolution for Multimodal 3D Object Detection. CVPR 2023.
13	LumiNet		code	96.27 %	99.23 %	88.94 %	0.1 s	1 core @ 2.5 Ghz (Python)

14	LPFusion			96.20 %	96.73 %	93.33 %	0.1 s	1 core @ 2.5 Ghz (C/C++)

15	GraR-Po		code	96.18 %	96.84 %	91.11 %	0.06 s	1 core @ 2.5 Ghz (Python + C/C++)
H. Yang, Z. Liu, X. Wu, W. Wang, W. Qian, X. He and D. Cai: Graph R-CNN: Towards Accurate 3D Object Detection with Semantic-Decorated Local Graph. ECCV 2022.
16	SFD		code	96.17 %	98.97 %	91.13 %	0.1 s	1 core @ 2.5 Ghz (Python + C/C++)
X. Wu, L. Peng, H. Yang, L. Xie, C. Huang, C. Deng, H. Liu and D. Cai: Sparse Fuse Dense: Towards High Quality 3D Detection with Depth Completion. CVPR 2022.
17	MLF-DET			96.17 %	96.89 %	88.90 %	0.09 s	1 core @ 2.5 Ghz (C/C++)
Z. Lin, Y. Shen, S. Zhou, S. Chen and N. Zheng: MLF-DET: Multi-Level Fusion for Cross- Modal 3D Object Detection. International Conference on Artificial Neural Networks 2023.
18	VPFNet		code	96.15 %	96.64 %	91.14 %	0.06 s	2 cores @ 2.5 Ghz (Python)
H. Zhu, J. Deng, Y. Zhang, J. Ji, Q. Mao, H. Li and Y. Zhang: VPFNet: Improving 3D Object Detection with Virtual Point based LiDAR and Stereo Data Fusion. IEEE Transactions on Multimedia 2022.
19	LPFusion_three_class			96.07 %	98.74 %	93.27 %	0.1 s	1 core @ 2.5 Ghz (C/C++)

20	CLOCs		code	96.07 %	96.77 %	91.11 %	0.1 s	1 core @ 2.5 Ghz (Python)
S. Pang, D. Morris and H. Radha: CLOCs: Camera-LiDAR Object Candidates Fusion for 3D Object Detection . 2020 IEEE/RSJ International Conference on Intelligent Robots and Systems (IROS) 2020.
21	ACFNet			96.06 %	96.68 %	93.36 %	0.11 s	1 core @ 2.5 Ghz (C/C++)
Y. Tian, X. Zhang, X. Wang, J. Xu, J. Wang, R. Ai, W. Gu and W. Ding: ACF-Net: Asymmetric Cascade Fusion for 3D Detection With LiDAR Point Clouds and Images. IEEE Transactions on Intelligent Vehicles 2023.
22	RDIoU		code	96.05 %	98.79 %	91.03 %	0.03 s	1 core @ 2.5 Ghz (Python + C/C++)
H. Sheng, S. Cai, N. Zhao, B. Deng, J. Huang, X. Hua, M. Zhao and G. Lee: Rethinking IoU-based Optimization for Single- stage 3D Object Detection. ECCV 2022.
23	GraR-Vo		code	96.05 %	96.67 %	93.01 %	0.04 s	1 core @ 2.5 Ghz (C/C++)
H. Yang, Z. Liu, X. Wu, W. Wang, W. Qian, X. He and D. Cai: Graph R-CNN: Towards Accurate 3D Object Detection with Semantic-Decorated Local Graph. ECCV 2022.
24	SFD++			96.03 %	98.91 %	93.23 %	0.06 s	1 core @ 2.5 Ghz (C/C++)

25	TED		code	96.03 %	96.64 %	93.35 %	0.1 s	1 core @ 2.5 Ghz (C/C++)
H. Wu, C. Wen, W. Li, R. Yang and C. Wang: Transformation-Equivariant 3D Object Detection for Autonomous Driving. AAAI 2023.
26	BFT3D			96.03 %	96.98 %	88.81 %	0.15 s	1 core @ 2.5 Ghz (C/C++)

27	CLOCs_PVCas		code	95.96 %	96.76 %	91.08 %	0.1 s	1 core @ 2.5 Ghz (Python)
S. Pang, D. Morris and H. Radha: CLOCs: Camera-LiDAR Object Candidates Fusion for 3D Object Detection . 2020 IEEE/RSJ International Conference on Intelligent Robots and Systems (IROS) 2020.
28	PVT-SSD			95.90 %	96.75 %	90.69 %	0.05 s	1 core @ 2.5 Ghz (Python + C/C++)
H. Yang, W. Wang, M. Chen, B. Lin, T. He, H. Chen, X. He and W. Ouyang: PVT-SSD: Single-Stage 3D Object Detector with Point-Voxel Transformer. CVPR 2023.
29	UPIDet		code	95.89 %	96.25 %	93.25 %	0.11 s	1 core @ 2.5 Ghz (C/C++)
Y. Zhang, Q. Zhang, J. Hou, Y. Yuan and G. Xing: Unleash the Potential of Image Branch for Cross-modal 3D Object Detection. Thirty-seventh Conference on Neural Information Processing Systems 2023.
30	GraR-Pi		code	95.89 %	98.59 %	92.85 %	0.03 s	1 core @ 2.5 Ghz (C/C++)
H. Yang, Z. Liu, X. Wu, W. Wang, W. Qian, X. He and D. Cai: Graph R-CNN: Towards Accurate 3D Object Detection with Semantic-Decorated Local Graph. ECCV 2022.
31	MPCF		code	95.87 %	98.95 %	90.98 %	0.08 s	1 core @ 2.5 Ghz (C/C++)
P. Gao and P. Zhang: MPCF: Multi-Phase Consolidated Fusion for Multi-Modal 3D Object Detection with Pseudo Point Cloud. 2024.
32	OcTr			95.84 %	96.48 %	90.99 %	0.06 s	GPU @ 2.5 Ghz (Python + C/C++)
C. Zhou, Y. Zhang, J. Chen and D. Huang: OcTr: Octree-based Transformer for 3D Object Detection. CVPR 2023.
33	DDF			95.83 %	96.53 %	93.25 %	0.1 s	1 core @ 2.5 Ghz (Python)

34	3D Dual-Fusion		code	95.82 %	96.54 %	93.11 %	0.1 s	1 core @ 2.5 Ghz (Python)
Y. Kim, K. Park, M. Kim, D. Kum and J. Choi: 3D Dual-Fusion: Dual-Domain Dual-Query Camera-LiDAR Fusion for 3D Object Detection. arXiv preprint arXiv:2211.13529 2022.
35	GLENet-VR		code	95.81 %	96.85 %	90.91 %	0.04 s	1 core @ 2.5 Ghz (C/C++)
Y. Zhang, Q. Zhang, Z. Zhu, J. Hou and Y. Yuan: GLENet: Boosting 3D object detectors with generative label uncertainty estimation. International Journal of Computer Vision 2023. Y. Zhang, J. Hou and Y. Yuan: A Comprehensive Study of the Robustness for LiDAR-based 3D Object Detectors against Adversarial Attacks. International Journal of Computer Vision 2023.
36	mm3d			95.81 %	96.89 %	90.92 %	0.2 s	1 core @ 2.5 Ghz (C/C++)

37	TSSTDet			95.81 %	96.65 %	93.05 %	0.08 s	1 core @ 2.5 Ghz (C/C++)
H. Hoang, D. Bui and M. Yoo: TSSTDet: Transformation-Based 3-D Object Detection via a Spatial Shape Transformer. IEEE Sensors Journal 2024.
38	R2Pfusion-Det			95.79 %	96.53 %	93.20 %	0.3 s	1 core @ 2.5 Ghz (C/C++)

39	DVF-V			95.77 %	96.60 %	90.89 %	0.1 s	1 core @ 2.5 Ghz (C/C++)
A. Mahmoud, J. Hu and S. Waslander: Dense Voxel Fusion for 3D Object Detection. WACV 2023.
40	Fast-CLOCs			95.75 %	96.69 %	90.95 %	0.1 s	GPU @ 2.5 Ghz (Python)
S. Pang, D. Morris and H. Radha: Fast-CLOCs: Fast Camera-LiDAR Object Candidates Fusion for 3D Object Detection. Proceedings of the IEEE/CVF Winter Conference on Applications of Computer Vision (WACV) 2022.
41	TRTConv-L			95.73 %	96.58 %	92.97 %	0.01 s	1 core @ 2.5 Ghz (C/C++)
ERROR: Wrong syntax in BIBTEX file.
42	3D HANet		code	95.73 %	98.61 %	92.96 %	0.1 s	1 core @ 2.5 Ghz (C/C++)
Q. Xia, Y. Chen, G. Cai, G. Chen, D. Xie, J. Su and Z. Wang: 3D HANet: A Flexible 3D Heatmap Auxiliary Network for Object Detection. IEEE Transactions on Geoscience and Remote Sensing 2023.
43	DSGN++		code	95.70 %	98.08 %	88.27 %	0.2 s	GeForce RTX 2080Ti
Y. Chen, S. Huang, S. Liu, B. Yu and J. Jia: DSGN++: Exploiting Visual-Spatial Relation for Stereo-Based 3D Detectors. IEEE Transactions on Pattern Analysis and Machine Intelligence 2022.
44	LCANet			95.66 %	96.56 %	93.06 %	1 s	1 core @ 2.5 Ghz (C/C++)

45	mat3D			95.64 %	98.83 %	93.00 %	0.2 s	1 core @ 2.5 Ghz (C/C++)

46	CasA		code	95.62 %	96.52 %	92.86 %	0.1 s	1 core @ 2.5 Ghz (C/C++)
H. Wu, J. Deng, C. Wen, X. Li and C. Wang: CasA: A Cascade Attention Network for 3D Object Detection from LiDAR point clouds. IEEE Transactions on Geoscience and Remote Sensing 2022.
47	BADet		code	95.61 %	98.75 %	90.64 %	0.14 s	1 core @ 2.5 Ghz (C/C++)
R. Qian, X. Lai and X. Li: BADet: Boundary-Aware 3D Object Detection from Point Clouds. Pattern Recognition 2022.
48	SE-SSD		code	95.60 %	96.69 %	90.53 %	0.03 s	1 core @ 2.5 Ghz (Python + C/C++)
W. Zheng, W. Tang, L. Jiang and C. Fu: SE-SSD: Self-Ensembling Single-Stage Object Detector From Point Cloud. CVPR 2021.
49	FARP-Net		code	95.57 %	96.11 %	93.07 %	0.06 s	GPU @ 2.5 Ghz (Python + C/C++)
T. Xie, L. Wang, K. Wang, R. Li, X. Zhang, H. Zhang, L. Yang, H. Liu and J. Li: FARP-Net: Local-Global Feature Aggregation and Relation-Aware Proposals for 3D Object Detection. IEEE Transactions on Multimedia 2023.
50	LoGoNet		code	95.55 %	96.60 %	93.07 %	0.1 s	1 core @ 2.5 Ghz (C/C++)
X. Li, T. Ma, Y. Hou, B. Shi, Y. Yang, Y. Liu, X. Wu, Q. Chen, Y. Li, Y. Qiao and others: LoGoNet: Towards Accurate 3D Object Detection with Local-to-Global Cross-Modal Fusion. CVPR 2023.
51	GD-MAE			95.54 %	98.38 %	90.42 %	0.07 s	1 core @ 2.5 Ghz (Python + C/C++)
H. Yang, T. He, J. Liu, H. Chen, B. Wu, B. Lin, X. He and W. Ouyang: GD-MAE: Generative Decoder for MAE Pre- training on LiDAR Point Clouds. CVPR 2023.
52	ECA			95.50 %	98.11 %	92.79 %	0.08 s	GPU @ 1.5 Ghz (Python)

53	DVF-PV			95.49 %	96.42 %	92.57 %	0.1 s	1 core @ 2.5 Ghz (C/C++)
A. Mahmoud, J. Hu and S. Waslander: Dense Voxel Fusion for 3D Object Detection. WACV 2023.
54	BVPConv-T			95.48 %	96.32 %	92.91 %	0.05 s	1 core @ 2.5 Ghz (Python + C/C++)
ERROR: Wrong syntax in BIBTEX file.
55	HS-fusion			95.48 %	98.33 %	92.66 %	- s	1 core @ 2.5 Ghz (Python)
ERROR: Wrong syntax in BIBTEX file.
56	SPANet			95.46 %	96.54 %	90.47 %	0.06 s	1 core @ 2.5 Ghz (C/C++)
Y. Ye: SPANet: Spatial and Part-Aware Aggregation Network for 3D Object Detection. Pacific Rim International Conference on Artificial Intelligence 2021.
57	PG-RCNN		code	95.40 %	96.66 %	90.55 %	0.06 s	GPU @ 1.5 Ghz (Python)
I. Koo, I. Lee, S. Kim, H. Kim, W. Jeon and C. Kim: PG-RCNN: Semantic Surface Point Generation for 3D Object Detection. 2023.
58	c2f			95.39 %	96.25 %	92.65 %	1 s	1 core @ 2.5 Ghz (C/C++)

59	FIRM-Net_SCF+			95.38 %	96.31 %	92.71 %	0.07 s	1 core @ 2.5 Ghz (Python)

60	SCDA-Net			95.37 %	98.62 %	92.90 %	- s	1 core @ 2.5 Ghz (C/C++)

61	FIRM-Net-SCF			95.36 %	96.30 %	92.69 %	0.07 s	1 core @ 2.5 Ghz (Python)

62	SASA		code	95.35 %	96.01 %	92.53 %	0.04 s	1 core @ 2.5 Ghz (Python + C/C++)
C. Chen, Z. Chen, J. Zhang and D. Tao: SASA: Semantics-Augmented Set Abstraction for Point-based 3D Object Detection. arXiv preprint arXiv:2201.01976 2022.
63	SPG_mini		code	95.32 %	96.23 %	92.68 %	0.09 s	GPU @ 2.5 Ghz (Python)
Q. Xu, Y. Zhou, W. Wang, C. Qi and D. Anguelov: SPG: Unsupervised Domain Adaptation for 3D Object Detection via Semantic Point Generation. Proceedings of the IEEE conference on computer vision and pattern recognition (ICCV) 2021.
64	EQ-PVRCNN		code	95.32 %	98.23 %	92.65 %	0.2 s	GPU @ 2.5 Ghz (Python + C/C++)
Z. Yang, L. Jiang, Y. Sun, B. Schiele and J. Jia: A Unified Query-based Paradigm for Point Cloud Understanding. Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition 2022.
65	TRTConv-T			95.31 %	98.38 %	92.69 %	0.04 s	1 core @ 2.5 Ghz (Python + C/C++)
ERROR: Wrong syntax in BIBTEX file.
66	BVPConv-L			95.30 %	96.37 %	92.76 %	0.01 s	1 core @ 2.5 Ghz (Python + C/C++)
ERROR: Wrong syntax in BIBTEX file.
67	Focals Conv		code	95.28 %	96.30 %	92.69 %	0.1 s	1 core @ 2.5 Ghz (C/C++)
Y. Chen, Y. Li, X. Zhang, J. Sun and J. Jia: Focal Sparse Convolutional Networks for 3D Object Detection. Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition 2022.
68	CasA++		code	95.28 %	95.83 %	94.28 %	0.1 s	1 core @ 2.5 Ghz (C/C++)
H. Wu, J. Deng, C. Wen, X. Li and C. Wang: CasA: A Cascade Attention Network for 3D Object Detection from LiDAR point clouds. IEEE Transactions on Geoscience and Remote Sensing 2022.
69	VoxSeT		code	95.23 %	96.16 %	90.49 %	33 ms	1 core @ 2.5 Ghz (C/C++)
C. He, R. Li, S. Li and L. Zhang: Voxel Set Transformer: A Set-to-Set Approach to 3D Object Detection from Point Clouds. CVPR 2022.
70	voxel-rcnn+++		code	95.22 %	96.40 %	92.37 %	0.08 s	GPU @ 2.5 Ghz (Python)

71	PC-CNN-V2			95.20 %	96.06 %	89.37 %	0.5 s	GPU @ 2.5 Ghz (Matlab + C/C++)
X. Du, M. Ang, S. Karaman and D. Rus: A General Pipeline for 3D Detection of Vehicles. 2018 IEEE International Conference on Robotics and Automation (ICRA) 2018.
72	RagNet3D		code	95.17 %	96.27 %	92.66 %	0.05 s	1 core @ 2.5 Ghz (C/C++)
J. Chen, Y. Han, Z. Yan, J. Qian, J. Li and J. Yang: Ragnet3d: Learning Distinguishable Representation for Pooled Grids in 3d Object Detection. Available at SSRN 4979473 .
73	VPFNet		code	95.17 %	96.06 %	92.66 %	0.2 s	1 core @ 2.5 Ghz (C/C++)
C. Wang, H. Chen and L. Fu: VPFNet: Voxel-Pixel Fusion Network for Multi-class 3D Object Detection. 2021. C. Wang, H. Chen, Y. Chen, P. Hsiao and L. Fu: VoPiFNet: Voxel-Pixel Fusion Network for Multi-Class 3D Object Detection. IEEE Transactions on Intelligent Transportation Systems 2024.
74	F-PointNet		code	95.17 %	95.85 %	85.42 %	0.17 s	GPU @ 3.0 Ghz (Python)
C. Qi, W. Liu, C. Wu, H. Su and L. Guibas: Frustum PointNets for 3D Object Detection from RGB-D Data. arXiv preprint arXiv:1711.08488 2017.
75	EPNet++			95.17 %	96.73 %	92.10 %	0.1 s	GPU @ 2.5 Ghz (Python)
Z. Liu, T. Huang, B. Li, X. Chen, X. Wang and X. Bai: EPNet++: Cascade Bi-Directional Fusion for Multi-Modal 3D Object Detection. IEEE Transactions on Pattern Analysis and Machine Intelligence 2022.
76	SA-SSD		code	95.16 %	97.92 %	90.15 %	0.04 s	1 core @ 2.5 Ghz (Python)
C. He, H. Zeng, J. Huang, X. Hua and L. Zhang: Structure Aware Single-stage 3D Object Detection from Point Cloud. CVPR 2020.
77	HMFI		code	95.16 %	96.29 %	92.45 %	0.1 s	1 core @ 2.5 Ghz (C/C++)
X. Li, B. Shi, Y. Hou, X. Wu, T. Ma, Y. Li and L. He: Homogeneous Multi-modal Feature Fusion and Interaction for 3D Object Detection. ECCV 2022.
78	USVLab BSAODet		code	95.15 %	96.26 %	92.62 %	0.04 s	1 core @ 2.5 Ghz (C/C++)
W. Xiao, Y. Peng, C. Liu, J. Gao, Y. Wu and X. Li: Balanced Sample Assignment and Objective for Single-Model Multi-Class 3D Object Detection. IEEE Transactions on Circuits and Systems for Video Technology 2023.
79	MPC3DNet			95.14 %	98.49 %	92.42 %	0.05 s	GPU @ 1.5 Ghz (Python)

80	Pyramid R-CNN			95.13 %	95.88 %	92.62 %	0.07 s	1 core @ 2.5 Ghz (C/C++)
J. Mao, M. Niu, H. Bai, X. Liang, H. Xu and C. Xu: Pyramid R-CNN: Towards Better Performance and Adaptability for 3D Object Detection. ICCV 2021.
81	Voxel R-CNN		code	95.11 %	96.49 %	92.45 %	0.04 s	GPU @ 3.0 Ghz (C/C++)
J. Deng, S. Shi, P. Li, W. Zhou, Y. Zhang and H. Li: Voxel R-CNN: Towards High Performance Voxel-based 3D Object Detection . AAAI 2021.
82	3DSSD		code	95.10 %	97.69 %	92.18 %	0.04 s	GPU @ 2.5 Ghz (Python + C/C++)
Z. Yang, Y. Sun, S. Liu and J. Jia: 3DSSD: Point-based 3D Single Stage Object Detector. CVPR 2020.
83	CAIA_PRO		code	95.09 %	95.72 %	90.44 %	0.01 s	1 core @ 2.5 Ghz (C/C++)

84	MonoSample (DID-M3D)		code	95.02 %	96.45 %	85.58 %	0.2 s	1 core @ 2.5 Ghz (C/C++)
J. Qiao, B. Liu, J. Yang, B. Wang, S. Xiu, X. Du and X. Nie: MonoSample: Synthetic 3D Data Augmentation Method in Monocular 3D Object Detection. IEEE Robotics and Automation Letters 2024.
85	PDV		code	95.00 %	96.07 %	92.44 %	0.1 s	1 core @ 2.5 Ghz (C/C++)
J. Hu, T. Kuai and S. Waslander: Point Density-Aware Voxels for LiDAR 3D Object Detection. CVPR 2022.
86	MVRA + I-FRCNN+			94.98 %	95.87 %	82.52 %	0.18 s	GPU @ 2.5 Ghz (Python)
H. Choi, H. Kang and Y. Hyun: Multi-View Reprojection Architecture for Orientation Estimation. The IEEE International Conference on Computer Vision (ICCV) Workshops 2019.
87	SVFMamba		code	94.97 %	95.54 %	92.24 %	N/A s	1 core @ 2.5 Ghz (C/C++)

88	SIENet		code	94.97 %	96.02 %	92.40 %	0.08 s	1 core @ 2.5 Ghz (Python)
Z. Li, Y. Yao, Z. Quan, W. Yang and J. Xie: SIENet: Spatial Information Enhancement Network for 3D Object Detection from Point Cloud. 2021.
89	VoTr-TSD		code	94.94 %	95.97 %	92.44 %	0.07 s	1 core @ 2.5 Ghz (C/C++)
J. Mao, Y. Xue, M. Niu, H. Bai, J. Feng, X. Liang, H. Xu and C. Xu: Voxel Transformer for 3D Object Detection. ICCV 2021.
90	L-AUG			94.92 %	95.84 %	92.22 %	0.1 s	1 core @ 2.5 Ghz (C/C++)
T. Cortinhal, I. Gouigah and E. Aksoy: Semantics-aware LiDAR-Only Pseudo Point Cloud Generation for 3D Object Detection. 2023.
91	SQD		code	94.92 %	98.21 %	92.37 %	0.06 s	1 core @ 2.5 Ghz (C/C++)
Y. Mo, Y. Wu, J. Zhao, Z. Hou, W. Huang, Y. Hu, J. Wang and J. Yan: Sparse Query Dense: Enhancing 3D Object Detection with Pseudo Points. ACM MM Oral 2024.
92	GraphAlign(ICCV2023)		code	94.87 %	98.06 %	92.47 %	0.03 s	GPU @ 2.0 Ghz (Python)
Z. Song, H. Wei, L. Bai, L. Yang and C. Jia: GraphAlign: Enhancing accurate feature alignment by graph matching for multi-modal 3D object detection. Proceedings of the IEEE/CVF International Conference on Computer Vision 2023.
93	M3DeTR		code	94.83 %	97.39 %	92.10 %	n/a s	GPU @ 1.0 Ghz (Python)
T. Guan, J. Wang, S. Lan, R. Chandra, Z. Wu, L. Davis and D. Manocha: M3DeTR: Multi-representation, Multi- scale, Mutual-relation 3D Object Detection with Transformers. 2021.
94	focal			94.83 %	95.91 %	92.48 %	100 s	1 core @ 2.5 Ghz (Python)

95	GEFPN			94.83 %	95.91 %	92.48 %	0.5 s	1 core @ 2.5 Ghz (C/C++)

96	GeVo			94.83 %	95.91 %	92.48 %	0.05 s	1 core @ 2.5 Ghz (C/C++)

97	StructuralIF			94.81 %	96.14 %	92.12 %	0.02 s	8 cores @ 2.5 Ghz (Python)
J. Pei An: Deep structural information fusion for 3D object detection on LiDAR-camera system. Accepted in CVIU 2021.
98	MLFusion-VS			94.79 %	98.26 %	92.43 %	0.06 s	1 core @ 2.5 Ghz (C/C++)

99	AFFN-G			94.78 %	98.07 %	92.32 %	0.5 s	1 core @ 2.5 Ghz (C/C++)

100	XView			94.77 %	95.89 %	92.23 %	0.1 s	1 core @ 2.5 Ghz (C/C++)
L. Xie, G. Xu, D. Cai and X. He: X-view: Non-egocentric Multi-View 3D Object Detector. 2021.
101	AFFN			94.75 %	95.91 %	92.16 %	0.5 s	1 core @ 2.5 Ghz (C/C++)

102	P2V-RCNN			94.73 %	96.03 %	92.34 %	0.1 s	1 core @ 2.5 Ghz (C/C++)
J. Li, S. Luo, Z. Zhu, H. Dai, A. Krylov, Y. Ding and L. Shao: P2V-RCNN: Point to Voxel Feature Learning for 3D Object Detection from Point Clouds. IEEE Access 2021.
103	SP_SECOND_IOU		code	94.72 %	95.85 %	92.19 %	0.04 s	1 core @ 2.5 Ghz (Python)

104	SPG		code	94.71 %	97.80 %	92.19 %	0.09 s	1 core @ 2.5 Ghz (Python + C/C++)
Q. Xu, Y. Zhou, W. Wang, C. Qi and D. Anguelov: SPG: Unsupervised Domain Adaptation for 3D Object Detection via Semantic Point Generation. Proceedings of the IEEE conference on computer vision and pattern recognition (ICCV) 2021.
105	CAT-Det			94.71 %	95.97 %	92.07 %	0.3 s	GPU @ 2.5 Ghz (Python + C/C++)
Y. Zhang, J. Chen and D. Huang: CAT-Det: Contrastively Augmented Transformer for Multi-modal 3D Object Detection. CVPR 2022.
106	test			94.70 %	97.49 %	91.88 %	0.04 s	GPU @ 1.5 Ghz (Python + C/C++)

107	MMLab PV-RCNN		code	94.70 %	98.17 %	92.04 %	0.08 s	1 core @ 2.5 Ghz (Python + C/C++)
S. Shi, C. Guo, L. Jiang, Z. Wang, J. Shi, X. Wang and H. Li: PV-RCNN: Point-Voxel Feature Set Abstraction for 3D Object Detection. CVPR 2020.
108	SVGA-Net			94.67 %	96.05 %	91.86 %	0.03s	1 core @ 2.5 Ghz (Python + C/C++)
Q. He, Z. Wang, H. Zeng, Y. Zeng and Y. Liu: SVGA-Net: Sparse Voxel-Graph Attention Network for 3D Object Detection from Point Clouds. AAAI 2022.
109	RangeDet (Official)		code	94.64 %	95.50 %	91.77 %	0.02 s	1 core @ 2.5 Ghz (C/C++)
L. Fan, X. Xiong, F. Wang, N. Wang and Z. Zhang: RangeDet: In Defense of Range View for LiDAR-Based 3D Object Detection. Proceedings of the IEEE/CVF International Conference on Computer Vision (ICCV) 2021.
110	DSA-PV-RCNN		code	94.64 %	95.86 %	92.10 %	0.08 s	1 core @ 2.5 Ghz (Python + C/C++)
P. Bhattacharyya, C. Huang and K. Czarnecki: SA-Det3D: Self-Attention Based Context-Aware 3D Object Detection. 2021.
111	CG-SSD			94.63 %	95.97 %	92.08 %	0.01 s	1 core @ 2.5 Ghz (C/C++)

112	second_iou_baseline		code	94.62 %	95.74 %	91.97 %	0.05 s	1 core @ 2.5 Ghz (Python)

113	RangeIoUDet			94.61 %	95.74 %	91.98 %	0.02 s	GPU @ 2.5 Ghz (Python + C/C++)
Z. Liang, Z. Zhang, M. Zhang, X. Zhao and S. Pu: RangeIoUDet: Range Image Based Real-Time 3D Object Detector Optimized by Intersection Over Union. CVPR 2021.
114	BVIFusion+			94.61 %	95.81 %	91.93 %	0.09 s	1 core @ 2.5 Ghz (Python)

115	PASS-PV-RCNN-Plus			94.59 %	95.79 %	92.10 %	1 s	1 core @ 2.5 Ghz (Python)
Anonymous: Leveraging Anchor-based LiDAR 3D Object Detection via Point Assisted Sample Selection. will submit to computer vision conference/journal 2024.
116	second_iou_baseline			94.58 %	95.75 %	91.96 %	0.03 s	1 core @ 2.5 Ghz (Python)

117	DVFENet			94.57 %	95.35 %	91.77 %	0.05 s	1 core @ 2.5 Ghz (Python + C/C++)
Y. He, G. Xia, Y. Luo, L. Su, Z. Zhang, W. Li and P. Wang: DVFENet: Dual-branch Voxel Feature Extraction Network for 3D Object Detection. Neurocomputing 2021.
118	MFB3D			94.49 %	97.64 %	91.83 %	0.14 s	1 core @ 2.5 Ghz (Python)

119	HA-PillarNet			94.48 %	98.03 %	92.14 %	0.05 s	1 core @ 2.5 Ghz (Python)

120	TuSimple		code	94.47 %	95.12 %	86.45 %	1.6 s	GPU @ 2.5 Ghz (Python + C/C++)
F. Yang, W. Choi and Y. Lin: Exploit all the layers: Fast and accurate cnn object detector with scale dependent pooling and cascaded rejection classifiers. Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition 2016. K. He, X. Zhang, S. Ren and J. Sun: Deep residual learning for image recognition. Proceedings of the IEEE conference on computer vision and pattern recognition 2016.
121	EPNet		code	94.44 %	96.15 %	89.99 %	0.1 s	1 core @ 2.5 Ghz (Python + C/C++)
T. Huang, Z. Liu, X. Chen and X. Bai: EPNet: Enhancing Point Features with Image Semantics for 3D Object Detection. ECCV 2020.
122	SFA_IGCL_Focalsconv*		code	94.44 %	95.92 %	92.18 %	0.2 s	1 core @ 2.5 Ghz (C/C++)

123	AFFN-Ga			94.44 %	95.80 %	92.02 %	0.5 s	1 core @ 2.5 Ghz (C/C++)

124	SERCNN			94.42 %	96.33 %	89.96 %	0.1 s	1 core @ 2.5 Ghz (Python)
D. Zhou, J. Fang, X. Song, L. Liu, J. Yin, Y. Dai, H. Li and R. Yang: Joint 3D Instance Segmentation and Object Detection for Autonomous Driving. Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition 2020.
125	dsvd+vx			94.30 %	95.09 %	91.51 %	0.1 s	1 core @ 2.5 Ghz (C/C++)

126	UberATG-MMF			94.25 %	97.41 %	89.87 %	0.08 s	GPU @ 2.5 Ghz (Python)
M. Liang, B. Yang, Y. Chen, R. Hu and R. Urtasun: Multi-Task Multi-Sensor Fusion for 3D Object Detection. CVPR 2019.
127	SRDL			94.24 %	95.86 %	91.80 %	0.05 s	1 core @ 2.5 Ghz (Python + C/C++)
ERROR: Wrong syntax in BIBTEX file.
128	CGML			94.14 %	97.56 %	91.89 %	0.33 s	1 core @ 2.5 Ghz (C/C++)

129	Voxel RCNN-Focal*		code	94.14 %	95.62 %	91.99 %	0.2 s	1 core @ 2.5 Ghz (Python)

130	FocalsConv*			94.10 %	97.67 %	91.88 %	0.2 s	1 core @ 2.5 Ghz (C/C++)

131	HMNet			94.07 %	95.51 %	91.23 %	0.1 s	1 core @ 2.5 Ghz (C/C++)

132	Res3DNet			94.05 %	95.44 %	91.32 %	0.05 s	GPU @ 3.5 Ghz (Python)

133	RangeRCNN			94.03 %	95.48 %	91.74 %	0.06 s	GPU @ 2.5 Ghz (Python + C/C++)
Z. Liang, M. Zhang, Z. Zhang, X. Zhao and S. Pu: RangeRCNN: Towards Fast and Accurate 3D Object Detection with Range Image Representation. arXiv preprint arXiv:2009.00206 2020.
134	SA V1			94.02 %	94.86 %	91.16 %	0.5 s	GPU @ 2.5 Ghz (Python)

135	SFA-GCL(80)		code	94.01 %	96.93 %	91.12 %	0.04 s	1 core @ 2.5 Ghz (C/C++)

136	Faraway-Frustum		code	93.99 %	95.81 %	91.72 %	0.1 s	GPU @ 2.5 Ghz (Python)
H. Zhang, D. Yang, E. Yurtsever, K. Redmill and U. Ozguner: Faraway-frustum: Dealing with lidar sparsity for 3D object detection using fusion. 2021 IEEE International Intelligent Transportation Systems Conference (ITSC) 2021.
137	DD3D		code	93.99 %	94.69 %	89.37 %	n/a s	1 core @ 2.5 Ghz (C/C++)
D. Park, R. Ambrus, V. Guizilini, J. Li and A. Gaidon: Is Pseudo-Lidar needed for Monocular 3D Object detection?. IEEE/CVF International Conference on Computer Vision (ICCV) .
138	MSMA V1			93.98 %	95.21 %	90.94 %	0.5 s	GPU @ 2.5 Ghz (Python)

139	SFA-GCL		code	93.97 %	96.90 %	91.09 %	0.04 s	1 core @ 2.5 Ghz (C/C++)

140	SFA-GCL(80, k=4)		code	93.96 %	96.88 %	91.07 %	0.04 s	1 core @ 2.5 Ghz (C/C++)

141	SIF			93.95 %	95.51 %	91.57 %	0.1 s	1 core @ 2.5 Ghz (C/C++)
P. An: SIF. Submitted to CVIU 2021.
142	MGAF-3DSSD		code	93.87 %	94.45 %	86.37 %	0.1 s	1 core @ 2.5 Ghz (Python)
J. Li, H. Dai, L. Shao and Y. Ding: Anchor-free 3D Single Stage Detector with Mask-Guided Attention for Point Cloud. MM '21: The 29th ACM International Conference on Multimedia (ACM MM) 2021.
143	3ONet			93.87 %	96.97 %	88.84 %	0.1 s	1 core @ 2.5 Ghz (C/C++)
H. Hoang and M. Yoo: 3ONet: 3-D Detector for Occluded Object Under Obstructed Conditions. IEEE Sensors Journal 2023.
144	LPCG-Monoflex		code	93.86 %	96.90 %	83.94 %	0.03 s	1 core @ 2.5 Ghz (C/C++)
L. Peng, F. Liu, Z. Yu, S. Yan, D. Deng, Z. Yang, H. Liu and D. Cai: Lidar Point Cloud Guided Monocular 3D Object Detection. ECCV 2022.
145	MMLAB LIGA-Stereo		code	93.82 %	96.43 %	86.19 %	0.4 s	1 core @ 2.5 Ghz (Python + C/C++)
X. Guo, S. Shi, X. Wang and H. Li: LIGA-Stereo: Learning LiDAR Geometry Aware Representations for Stereo-based 3D Detector. Proceedings of the IEEE/CVF International Conference on Computer Vision (ICCV) 2021.
146	SFA-GCL(baseline)		code	93.79 %	96.84 %	90.88 %	0.04 s	1 core @ 2.5 Ghz (C/C++)

147	SFA-GCL_dataaug		code	93.78 %	96.75 %	90.85 %	0.04 s	1 core @ 2.5 Ghz (C/C++)

148	SFA-GCL		code	93.78 %	96.87 %	90.84 %	0.04 s	1 core @ 2.5 Ghz (C/C++)

149	Sem-Aug			93.77 %	96.79 %	88.78 %	0.1 s	GPU @ 2.5 Ghz (Python)
L. Zhao, M. Wang and Y. Yue: Sem-Aug: Improving Camera-LiDAR Feature Fusion With Semantic Augmentation for 3D Vehicle Detection. IEEE Robotics and Automation Letters 2022.
150	Patches - EMP			93.75 %	97.91 %	90.56 %	0.5 s	GPU @ 2.5 Ghz (Python)
J. Lehner, A. Mitterecker, T. Adler, M. Hofmarcher, B. Nessler and S. Hochreiter: Patch Refinement: Localized 3D Object Detection. arXiv preprint arXiv:1910.04093 2019.
151	CIA-SSD		code	93.72 %	96.87 %	86.20 %	0.03 s	1 core @ 2.5 Ghz (Python + C/C++)
W. Zheng, W. Tang, S. Chen, L. Jiang and C. Fu: CIA-SSD: Confident IoU-Aware Single-Stage Object Detector From Point Cloud. AAAI 2021.
152	QD-3DT		code	93.66 %	94.26 %	83.63 %	0.03 s	GPU @ 2.5 Ghz (Python)
H. Hu, Y. Yang, T. Fischer, F. Yu, T. Darrell and M. Sun: Monocular Quasi-Dense 3D Object Tracking. ArXiv:2103.07351 2021.
153	MVAF-Net		code	93.66 %	95.37 %	90.90 %	0.06 s	1 core @ 2.5 Ghz (Python + C/C++)
G. Wang, B. Tian, Y. Zhang, L. Chen, D. Cao and J. Wu: Multi-View Adaptive Fusion Network for 3D Object Detection. arXiv preprint arXiv:2011.00652 2020.
154	SSL-PointGNN		code	93.65 %	96.61 %	88.53 %	0.56 s	GPU @ 1.5 Ghz (Python)
E. Erçelik, E. Yurtsever, M. Liu, Z. Yang, H. Zhang, P. Topçam, M. Listl, Y. Çaylı and A. Knoll: 3D Object Detection with a Self-supervised Lidar Scene Flow Backbone. arXiv preprint arXiv:2205.00705 2022.
155	PA3DNet			93.62 %	96.57 %	88.65 %	0.1 s	GPU @ 2.5 Ghz (Python + C/C++)
M. Wang, L. Zhao and Y. Yue: PA3DNet: 3-D Vehicle Detection with Pseudo Shape Segmentation and Adaptive Camera- LiDAR Fusion. IEEE Transactions on Industrial Informatics 2023.
156	IA-SSD (multi)		code	93.56 %	96.10 %	90.68 %	0.014 s	1 core @ 2.5 Ghz (C/C++)
Y. Zhang, Q. Hu, G. Xu, Y. Ma, J. Wan and Y. Guo: Not All Points Are Equal: Learning Highly Efficient Point-based Detectors for 3D LiDAR Point Clouds. CVPR 2022.
157	MonoLiG		code	93.56 %	96.70 %	83.74 %	0.03 s	1 core @ 2.5 Ghz (C/C++)
A. Hekimoglu, M. Schmidt and A. Ramiro: Monocular 3D Object Detection with LiDAR Guided Semi Supervised Active Learning. 2023.
158	MonoPair			93.55 %	96.61 %	83.55 %	0.06 s	GPU @ 2.5 Ghz (Python + C/C++)
Y. Chen, L. Tai, K. Sun and M. Li: MonoPair: Monocular 3D Object Detection Using Pairwise Spatial Relationships. The IEEE Conference on Computer Vision and Pattern Recognition (CVPR) 2020.
159	SCNet3D			93.54 %	96.55 %	90.89 %	0.08 s	1 core @ 2.5 Ghz (Python)

160	IA-SSD (single)		code	93.54 %	96.26 %	88.49 %	0.013 s	1 core @ 2.5 Ghz (C/C++)
Y. Zhang, Q. Hu, G. Xu, Y. Ma, J. Wan and Y. Guo: Not All Points Are Equal: Learning Highly Efficient Point-based Detectors for 3D LiDAR Point Clouds. CVPR 2022.
161	EBM3DOD		code	93.54 %	96.81 %	88.33 %	0.12 s	1 core @ 2.5 Ghz (Python)
F. Gustafsson, M. Danelljan and T. Schön: Accurate 3D Object Detection using Energy- Based Models. arXiv preprint arXiv:2012.04634 2020.
162	SeSame-point		code	93.50 %	95.22 %	90.44 %	N/A s	TITAN RTX @ 1.35 Ghz (Python)
H. O, C. Yang and K. Huh: SeSame: Simple, Easy 3D Object Detection with Point-Wise Semantics. Proceedings of the Asian Conference on Computer Vision (ACCV) 2024.
163	Deep MANTA			93.50 %	98.89 %	83.21 %	0.7 s	GPU @ 2.5 Ghz (Python + C/C++)
F. Chabot, M. Chaouch, J. Rabarisoa, C. Teulière and T. Chateau: Deep MANTA: A Coarse-to-fine Many-Task Network for joint 2D and 3D vehicle analysis from monocular image. CVPR 2017.
164	Point-GNN		code	93.50 %	96.58 %	88.35 %	0.6 s	GPU @ 2.5 Ghz (Python)
W. Shi and R. Rajkumar: Point-GNN: Graph Neural Network for 3D Object Detection in a Point Cloud. CVPR 2020.
165	BtcDet		code	93.47 %	96.23 %	88.55 %	0.09 s	GPU @ 2.5 Ghz (Python + C/C++)
Q. Xu, Y. Zhong and U. Neumann: Behind the Curtain: Learning Occluded Shapes for 3D Object Detection. Proceedings of the AAAI Conference on Artificial Intelligence 2022.
166	Struc info fusion II			93.45 %	96.72 %	88.31 %	0.05 s	GPU @ 2.5 Ghz (Python)
P. An, J. Liang, J. Ma, K. Yu and B. Fang: Struc info fusion. Submitted to CVIU 2021.
167	EBM3DOD baseline		code	93.45 %	96.72 %	88.25 %	0.05 s	1 core @ 2.5 Ghz (Python)
F. Gustafsson, M. Danelljan and T. Schön: Accurate 3D Object Detection using Energy- Based Models. arXiv preprint arXiv:2012.04634 2020.
168	StereoDistill			93.43 %	97.61 %	87.71 %	0.4 s	1 core @ 2.5 Ghz (Python)
Z. Liu, X. Ye, X. Tan, D. Errui, Y. Zhou and X. Bai: StereoDistill: Pick the Cream from LiDAR for Distilling Stereo-based 3D Object Detection. Proceedings of the AAAI Conference on Artificial Intelligence 2023.
169	MonoAFKD			93.42 %	96.18 %	83.62 %	0.03 s	1 core @ 2.5 Ghz (C/C++)

170	MonoLSS			93.42 %	96.19 %	83.62 %	0.04 s	1 core @ 2.5 Ghz (Python)
Z. Li, J. Jia and Y. Shi: MonoLSS: Learnable Sample Selection For Monocular 3D Detection. International Conference on 3D Vision 2024.
171	RRC		code	93.40 %	95.68 %	87.37 %	3.6 s	GPU @ 2.5 Ghz (C/C++)
J. Ren, X. Chen, J. Liu, W. Sun, J. Pang, Q. Yan, Y. Tai and L. Xu: Accurate Single Stage Detector Using Recurrent Rolling Convolution. CVPR 2017.
172	AM			93.39 %	96.22 %	85.84 %	0.04 s	1 core @ 2.5 Ghz (C/C++)

173	GSG-FPS		code	93.38 %	95.73 %	90.49 %	0.01 s	1 core @ 2.5 Ghz (C/C++)

174	3D-CVF at SPA		code	93.36 %	96.78 %	86.11 %	0.06 s	1 core @ 2.5 Ghz (C/C++)
J. Yoo, Y. Kim, J. Kim and J. Choi: 3D-CVF: Generating Joint Camera and LiDAR Features Using Cross-View Spatial Feature Fusion for 3D Object Detection. ECCV 2020.
175	SNVC		code	93.32 %	96.33 %	85.81 %	1 s	GPU @ 1.0 Ghz (Python)
S. Li, Z. Liu, Z. Shen and K. Cheng: Stereo Neural Vernier Caliper. Proceedings of the AAAI Conference on Artificial Intelligence 2022.
176	DFAF3D			93.32 %	96.58 %	90.24 %	0.05 s	1 core @ 2.5 Ghz (Python)
Q. Tang, X. Bai, J. Guo, B. Pan and W. Jiang: DFAF3D: A dual-feature-aware anchor-free single-stage 3D detector for point clouds. Image and Vision Computing 2023.
177	Struc info fusion I			93.31 %	96.59 %	88.23 %	0.05 s	1 core @ 2.5 Ghz (Python)
P. An, J. Liang, J. Ma, K. Yu and B. Fang: Struc info fusion. Submitted to CVIU 2021.
178	CityBrainLab-CT3D		code	93.30 %	96.28 %	90.58 %	0.07 s	1 core @ 2.5 Ghz (Python + C/C++)
H. Sheng, S. Cai, Y. Liu, B. Deng, J. Huang, X. Hua and M. Zhao: Improving 3D Object Detection with Channel- wise Transformer. ICCV 2021.
179	Test_dif		code	93.24 %	95.78 %	90.40 %	0.01 s	1 core @ 2.5 Ghz (C/C++)

180	STD		code	93.22 %	96.14 %	90.53 %	0.08 s	GPU @ 2.5 Ghz (Python + C/C++)
Z. Yang, Y. Sun, S. Liu, X. Shen and J. Jia: STD: Sparse-to-Dense 3D Object Detector for Point Cloud. ICCV 2019.
181	SARPNET			93.21 %	96.07 %	88.09 %	0.05 s	1 core @ 2.5 Ghz (Python + C/C++)
Y. Ye, H. Chen, C. Zhang, X. Hao and Z. Zhang: SARPNET: Shape Attention Regional Proposal Network for LiDAR-based 3D Object Detection. Neurocomputing 2019.
182	R50_SACINet			93.20 %	95.84 %	90.38 %	0.06 s	1 core @ 2.5 Ghz (C/C++)

183	H^23D R-CNN		code	93.20 %	96.20 %	90.55 %	0.03 s	1 core @ 2.5 Ghz (C/C++)
J. Deng, W. Zhou, Y. Zhang and H. Li: From Multi-View to Hollow-3D: Hallucinated Hollow-3D R-CNN for 3D Object Detection. IEEE Transactions on Circuits and Systems for Video Technology 2021.
184	Fast Point R-CNN			93.18 %	96.13 %	87.68 %	0.06 s	GPU @ 2.5 Ghz (Python + C/C++)
Y. Chen, S. Liu, X. Shen and J. Jia: Fast Point R-CNN. Proceedings of the IEEE international conference on computer vision (ICCV) 2019.
185	sensekitti		code	93.17 %	94.79 %	84.38 %	4.5 s	GPU @ 2.5 Ghz (Python + C/C++)
B. Yang, J. Yan, Z. Lei and S. Li: Craft Objects from Images. CVPR 2016.
186	SJTU-HW			93.11 %	96.30 %	82.21 %	0.85s	GPU @ 1.5 Ghz (Python + C/C++)
S. Zhang, X. Zhao, L. Fang, F. Haiping and S. Haitao: LED: LOCALIZATION-QUALITY ESTIMATION EMBEDDED DETECTOR. IEEE International Conference on Image Processing 2018. L. Fang, X. Zhao and S. Zhang: Small-objectness sensitive detection based on shifted single shot detector. Multimedia Tools and Applications 2018.
187	FromVoxelToPoint		code	93.06 %	96.08 %	90.53 %	0.1 s	1 core @ 2.5 Ghz (C/C++)
J. Li, H. Dai, L. Shao and Y. Ding: From Voxel to Point: IoU-guided 3D Object Detection for Point Cloud with Voxel-to- Point Decoder. MM '21: The 29th ACM International Conference on Multimedia (ACM MM) 2021.
188	L_SACINet			93.01 %	95.66 %	88.18 %	0.04 s	1 core @ 2.5 Ghz (C/C++)

189	CLOCs_SecCas			92.95 %	95.43 %	89.21 %	0.1 s	1 core @ 2.5 Ghz (Python)
S. Pang, D. Morris and H. Radha: CLOCs: Camera-LiDAR Object Candidates Fusion for 3D Object Detection. 2020 IEEE/RSJ International Conference on Intelligent Robots and Systems (IROS) 2020.
190	MonoCD		code	92.91 %	96.43 %	85.55 %	n/a s	1 core @ 2.5 Ghz (Python)
L. Yan, P. Yan, S. Xiong, X. Xiang and Y. Tan: MonoCD: Monocular 3D Object Detection with Complementary Depths. CVPR 2024.
191	Fade-kd			92.91 %	96.26 %	89.99 %	0.01 s	1 core @ 2.5 Ghz (C/C++)

192	ACDet		code	92.84 %	96.18 %	89.83 %	0.05 s	1 core @ 2.5 Ghz (C/C++)
J. Xu, G. Wang, X. Zhang and G. Wan: ACDet: Attentive Cross-view Fusion for LiDAR-based 3D Object Detection. 3DV 2022.
193	HotSpotNet			92.81 %	96.21 %	89.80 %	0.04 s	1 core @ 2.5 Ghz (Python + C/C++)
Q. Chen, L. Sun, Z. Wang, K. Jia and A. Yuille: object as hotspots. Proceedings of the European Conference on Computer Vision (ECCV) 2020.
194	SecAtten			92.81 %	95.51 %	89.78 %	0.1 s	1 core @ 2.5 Ghz (C/C++)

195	SegVoxelNet			92.73 %	96.00 %	87.60 %	0.04 s	1 core @ 2.5 Ghz (Python)
H. Yi, S. Shi, M. Ding, J. Sun, K. Xu, H. Zhou, Z. Wang, S. Li and G. Wang: SegVoxelNet: Exploring Semantic Context and Depth-aware Features for 3D Vehicle Detection from Point Cloud. ICRA 2020.
196	Patches			92.72 %	96.34 %	87.63 %	0.15 s	GPU @ 2.0 Ghz
J. Lehner, A. Mitterecker, T. Adler, M. Hofmarcher, B. Nessler and S. Hochreiter: Patch Refinement: Localized 3D Object Detection. arXiv preprint arXiv:1910.04093 2019.
197	Cube R-CNN		code	92.72 %	95.78 %	84.81 %	0.05 s	GPU @ 2.5 Ghz (Python)
G. Brazil, A. Kumar, J. Straub, N. Ravi, J. Johnson and G. Gkioxari: Omni3D: A Large Benchmark and Model for 3D Object Detection in the Wild. CVPR 2023.
198	CenterNet3D			92.69 %	95.76 %	89.81 %	0.04 s	GPU @ 1.5 Ghz (Python)
G. Wang, B. Tian, Y. Ai, T. Xu, L. Chen and D. Cao: CenterNet3D:An Anchor free Object Detector for Autonomous Driving. 2020.
199	R-GCN			92.67 %	96.19 %	87.66 %	0.16 s	GPU @ 2.5 Ghz (Python)
J. Zarzar, S. Giancola and B. Ghanem: PointRGCN: Graph Convolution Networks for 3D Vehicles Detection Refinement. ArXiv 2019.
200	zqd_test			92.66 %	95.89 %	87.10 %	0.2 s	1 core @ 2.5 Ghz (Python)

201	PI-RCNN			92.66 %	96.17 %	87.68 %	0.1 s	1 core @ 2.5 Ghz (Python)
L. Xie, C. Xiang, Z. Yu, G. Xu, Z. Yang, D. Cai and X. He: PI-RCNN: An Efficient Multi-sensor 3D Object Detector with Point-based Attentive Cont-conv Fusion Module. AAAI 2020 : The Thirty-Fourth AAAI Conference on Artificial Intelligence 2020.
202	EAEPNet			92.66 %	96.32 %	87.72 %	0.1 s	1 core @ 2.5 Ghz (Python)

203	PointPainting			92.58 %	98.39 %	89.71 %	0.4 s	GPU @ 2.5 Ghz (Python + C/C++)
S. Vora, A. Lang, B. Helou and O. Beijbom: PointPainting: Sequential Fusion for 3D Object Detection. CVPR 2020.
204	Fade			92.55 %	97.71 %	87.50 %	0.01 s	1 core @ 2.5 Ghz (C/C++)

205	DensePointPillars			92.55 %	95.66 %	87.35 %	0.03 s	GPU @ 2.5 Ghz (Python)

206	DASS			92.53 %	96.23 %	87.75 %	0.09 s	1 core @ 2.0 Ghz (Python)
O. Unal, L. Van Gool and D. Dai: Improving Point Cloud Semantic Segmentation by Learning 3D Object Detection. Proceedings of the IEEE/CVF Winter Conference on Applications of Computer Vision (WACV) 2021.
207	3D IoU-Net			92.47 %	96.31 %	87.67 %	0.1 s	1 core @ 2.5 Ghz (C/C++)
J. Li, S. Luo, Z. Zhu, H. Dai, S. Krylov, Y. Ding and L. Shao: 3D IoU-Net: IoU Guided 3D Object Detector for Point Clouds. arXiv preprint arXiv:2004.04962 2020.
208	Associate-3Ddet		code	92.45 %	95.61 %	87.32 %	0.05 s	1 core @ 2.5 Ghz (Python + C/C++)
L. Du, X. Ye, X. Tan, J. Feng, Z. Xu, E. Ding and S. Wen: Associate-3Ddet: Perceptual-to-Conceptual Association for 3D Point Cloud Object Detection. The IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) 2020.
209	S-AT GCN			92.44 %	95.06 %	90.78 %	0.02 s	GPU @ 2.0 Ghz (Python)
L. Wang, C. Wang, X. Zhang, T. Lan and J. Li: S-AT GCN: Spatial-Attention Graph Convolution Network based Feature Enhancement for 3D Object Detection. CoRR 2021.
210	PointRGCN			92.33 %	97.51 %	87.07 %	0.26 s	GPU @ V100 (Python)
J. Zarzar, S. Giancola and B. Ghanem: PointRGCN: Graph Convolution Networks for 3D Vehicles Detection Refinement. ArXiv 2019.
211	Sem-Aug-PointRCNN++			92.32 %	95.65 %	87.62 %	0.1 s	8 cores @ 3.0 Ghz (Python)
L. Zhao, M. Wang and Y. Yue: Sem-Aug: Improving Camera-LiDAR Feature Fusion With Semantic Augmentation for 3D Vehicle Detection. IEEE Robotics and Automation Letters 2022.
212	Harmonic PointPillar		code	92.25 %	95.16 %	89.11 %	0.01 s	1 core @ 2.5 Ghz (Python)
H. Zhang, J. Mekala, Z. Nain, J. Park and H. Jung: 3D Harmonic Loss: Towards Task-consistent and Time-friendly 3D Object Detection for V2X Orchestration. will submit to IEEE Transactions on Vehicular Technology 2022.
213	F-ConvNet		code	92.19 %	95.85 %	80.09 %	0.47 s	GPU @ 2.5 Ghz (Python + C/C++)
Z. Wang and K. Jia: Frustum ConvNet: Sliding Frustums to Aggregate Local Point-Wise Features for Amodal 3D Object Detection. IROS 2019.
214	PFF3D		code	92.15 %	95.37 %	87.54 %	0.05 s	GPU @ 3.0 Ghz (Python + C/C++)
L. Wen and K. Jo: Fast and Accurate 3D Object Detection for Lidar-Camera-Based Autonomous Vehicles Using One Shared Voxel-Based Backbone. IEEE Access 2021.
215	geo-pillars			92.10 %	95.30 %	89.08 %	0.05 s	1 core @ 2.5 Ghz (C/C++)
ERROR: Wrong syntax in BIBTEX file.
216	PASS-PointPillar			92.09 %	95.20 %	88.73 %	1 s	1 core @ 2.5 Ghz (C/C++)
Anonymous: Leveraging Anchor-based LiDAR 3D Object Detection via Point Assisted Sample Selection. will submit to computer vision conference/journal 2024.
217	SDP+RPN			92.03 %	95.16 %	79.16 %	0.4 s	GPU @ 2.5 Ghz (Python + C/C++)
F. Yang, W. Choi and Y. Lin: Exploit All the Layers: Fast and Accurate CNN Object Detector with Scale Dependent Pooling and Cascaded Rejection Classifiers. Proceedings of the IEEE International Conference on Computer Vision and Pattern Recognition 2016. S. Ren, K. He, R. Girshick and J. Sun: Faster R-CNN: Towards real-time object detection with region proposal networks. Advances in Neural Information Processing Systems 2015.
218	AEPF			92.02 %	95.48 %	87.50 %	0.05 s	GPU @ 2.5 Ghz (Python)

219	AB3DMOT		code	92.00 %	95.88 %	86.98 %	0.0047s	1 core @ 2.5 Ghz (Python)
X. Weng and K. Kitani: A Baseline for 3D Multi-Object Tracking. arXiv:1907.03961 2019.
220	PCNet3D++			91.90 %	94.96 %	88.61 %	0.5 s	GPU @ 3.5 Ghz (Python)

221	MMLab-PointRCNN		code	91.90 %	95.92 %	87.11 %	0.1 s	GPU @ 2.5 Ghz (Python + C/C++)
S. Shi, X. Wang and H. Li: Pointrcnn: 3d object proposal generation and detection from point cloud. Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition 2019.
222	MMLab-PartA^2		code	91.86 %	95.03 %	89.06 %	0.08 s	GPU @ 2.5 Ghz (Python + C/C++)
S. Shi, Z. Wang, J. Shi, X. Wang and H. Li: From Points to Parts: 3D Object Detection from Point Cloud with Part-aware and Part-aggregation Network. IEEE Transactions on Pattern Analysis and Machine Intelligence 2020.
223	mmFUSION		code	91.84 %	95.69 %	87.05 %	1s	1 core @ 2.5 Ghz (Python)
J. Ahmad and A. Del Bue: mmFUSION: Multimodal Fusion for 3D Objects Detection. arXiv preprint arXiv:2311.04058 2023.
224	WeakM3D		code	91.81 %	94.51 %	85.35 %	0.08 s	1 core @ 2.5 Ghz (C/C++)
L. Peng, S. Yan, B. Wu, Z. Yang, X. He and D. Cai: WeakM3D: Towards Weakly Supervised Monocular 3D Object Detection. ICLR 2022.
225	epBRM		code	91.77 %	94.59 %	88.45 %	0.1 s	GPU @ >3.5 Ghz (Python + C/C++)
K. Shin: Improving a Quality of 3D Object Detection by Spatial Transformation Mechanism. arXiv preprint arXiv:1910.04853 2019.
226	PL++: PV-RCNN++			91.77 %	94.79 %	88.82 %	0.342 s	RTX 4060Ti (Python)

227	PCNet3D			91.73 %	95.09 %	88.31 %	0.05 s	GPU @ 3.5 Ghz (Python)

228	C-GCN			91.73 %	95.64 %	86.37 %	0.147 s	GPU @ V100 (Python)
J. Zarzar, S. Giancola and B. Ghanem: PointRGCN: Graph Convolution Networks for 3D Vehicles Detection Refinement. ArXiv 2019.
229	ITVD		code	91.73 %	95.85 %	79.31 %	0.3 s	GPU @ 2.5 Ghz (C/C++)
Y. Wei Liu: Improving Tiny Vehicle Detection in Complex Scenes. IEEE International Conference on Multimedia and Expo (ICME) 2018.
230	BFT3D_easy			91.72 %	97.15 %	84.32 %	0.18 s	1 core @ 2.5 Ghz (C/C++)

231	SINet+		code	91.67 %	94.17 %	78.60 %	0.3 s	TITAN X GPU
X. Hu, X. Xu, Y. Xiao, H. Chen, S. He, J. Qin and P. Heng: SINet: A Scale-insensitive Convolutional Neural Network for Fast Vehicle Detection. IEEE Transactions on Intelligent Transportation Systems 2019.
232	Cascade MS-CNN		code	91.60 %	94.26 %	78.84 %	0.25 s	GPU @ 2.5 Ghz (C/C++)
Z. Cai and N. Vasconcelos: Cascade R-CNN: High Quality Object Detection and Instance Segmentation. arXiv preprint arXiv:1906.09756 2019. Z. Cai, Q. Fan, R. Feris and N. Vasconcelos: A unified multi-scale deep convolutional neural network for fast object detection. European conference on computer vision 2016.
233	SeSame-pillar		code	91.57 %	95.13 %	88.41 %	N/A s	TITAN RTX @ 1.35 Ghz (Python)
H. O, C. Yang and K. Huh: SeSame: Simple, Easy 3D Object Detection with Point-Wise Semantics. Proceedings of the Asian Conference on Computer Vision (ACCV) 2024.
234	PointRGBNet			91.48 %	95.40 %	86.50 %	0.08 s	4 cores @ 2.5 Ghz (Python + C/C++)
P. Xie Desheng: Real-time Detection of 3D Objects Based on Multi-Sensor Information Fusion. Automotive Engineering 2022.
235	MAFF-Net(DAF-Pillar)			91.46 %	94.38 %	83.89 %	0.04 s	1 core @ 2.5 Ghz (Python + C/C++)
Z. Zhang, Z. Liang, M. Zhang, X. Zhao, Y. Ming, T. Wenming and S. Pu: MAFF-Net: Filter False Positive for 3D Vehicle Detection with Multi-modal Adaptive Feature Fusion. arXiv preprint arXiv:2009.10945 2020.
236	zqd_test2			91.46 %	95.08 %	85.85 %	0.5 s	1 core @ 2.5 Ghz (C/C++)

237	HRI-VoxelFPN			91.44 %	96.65 %	86.18 %	0.02 s	GPU @ 2.5 Ghz (Python + C/C++)
H. Kuang, B. Wang, J. An, M. Zhang and Z. Zhang: Voxel-FPN:multi-scale voxel feature aggregation in 3D object detection from point clouds. sensors 2020.
238	EgoNet		code	91.39 %	96.18 %	81.33 %	0.1 s	GPU @ 1.5 Ghz (Python)
S. Li, Z. Yan, H. Li and K. Cheng: Exploring intermediate representation for monocular vehicle pose estimation. The IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) 2021.
239	Anonymous			91.37 %	94.04 %	88.23 %	0.02 s	GPU @ 2.5 Ghz (Python)

240	MonoDTF			91.35 %	95.03 %	85.92 %	0.1 s	1 core @ 2.5 Ghz (Python)

241	SeSame-pillar w/scor		code	91.34 %	94.89 %	88.13 %	N/A s	1 core @ 2.5 Ghz (C/C++)
H. O, C. Yang and K. Huh: SeSame: Simple, Easy 3D Object Detection with Point-Wise Semantics. Proceedings of the Asian Conference on Computer Vision (ACCV) 2024.
242	MonoSKD		code	91.34 %	96.68 %	83.69 %	0.04 s	1 core @ 2.5 Ghz (Python)
S. Wang and J. Zheng: MonoSKD: General Distillation Framework for Monocular 3D Object Detection via Spearman Correlation Coefficient. ECAI 2023.
243	zqd			91.30 %	94.88 %	85.58 %	0.1 s	1 core @ 2.5 Ghz (Python)

244	Stereo CenterNet			91.27 %	96.61 %	83.50 %	0.04 s	GPU @ 2.5 Ghz (Python)
Y. Shi, Y. Guo, Z. Mi and X. Li: Stereo CenterNet-based 3D object detection for autonomous driving. Neurocomputing 2022.
245	PointPillars		code	91.19 %	94.00 %	88.17 %	16 ms	1080ti GPU and Intel i7 CPU
A. Lang, S. Vora, H. Caesar, L. Zhou, J. Yang and O. Beijbom: PointPillars: Fast Encoders for Object Detection from Point Clouds. CVPR 2019.
246	LTN			91.18 %	94.68 %	81.51 %	0.4 s	GPU @ >3.5 Ghz (Python)
T. Wang, X. He, Y. Cai and G. Xiao: Learning a Layout Transfer Network for Context Aware Object Detection. IEEE Transactions on Intelligent Transportation Systems 2019.
247	MonoCoP-Car			91.18 %	93.60 %	81.77 %	0.01 s	GPU @ 2.5 Ghz (Python)

248	EOTL		code	91.17 %	96.31 %	81.20 %	TBD s	1 core @ 2.5 Ghz (Python + C/C++)
R. Yang, Z. Yan, T. Yang, Y. Wang and Y. Ruichek: Efficient Online Transfer Learning for Road Participants Detection in Autonomous Driving. IEEE Sensors Journal 2023.
249	WS3D			91.15 %	95.13 %	86.52 %	0.1 s	GPU @ 2.5 Ghz (Python)
Q. Meng, W. Wang, T. Zhou, J. Shen, L. Van Gool and D. Dai: Weakly Supervised 3D Object Detection from Lidar Point Cloud. 2020.
250	NeurOCS			91.08 %	96.39 %	81.20 %	0.1 s	GPU @ 2.5 Ghz (Python)

	Method	Setting	Code	Moderate	Easy	Hard	Runtime	Environment
1	F-PointNet		code	80.13 %	89.83 %	75.05 %	0.17 s	GPU @ 3.0 Ghz (Python)
C. Qi, W. Liu, C. Wu, H. Su and L. Guibas: Frustum PointNets for 3D Object Detection from RGB-D Data. arXiv preprint arXiv:1711.08488 2017.
2	HHA-TFFEM			78.53 %	87.01 %	74.70 %	0.14 s	GPU @ 2.5 Ghz (Python + C/C++)
F. Tan, Z. Xia, Y. Ma and X. Feng: 3D Sensor Based Pedestrian Detection by Integrating Improved HHA Encoding and Two-Branch Feature Fusion. Remote Sensing 2022.
3	TuSimple		code	78.40 %	88.87 %	73.66 %	1.6 s	GPU @ 2.5 Ghz (Python + C/C++)
F. Yang, W. Choi and Y. Lin: Exploit all the layers: Fast and accurate cnn object detector with scale dependent pooling and cascaded rejection classifiers. Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition 2016. K. He, X. Zhang, S. Ren and J. Sun: Deep residual learning for image recognition. Proceedings of the IEEE conference on computer vision and pattern recognition 2016.
4	RRC		code	76.61 %	85.98 %	71.47 %	3.6 s	GPU @ 2.5 Ghz (C/C++)
J. Ren, X. Chen, J. Liu, W. Sun, J. Pang, Q. Yan, Y. Tai and L. Xu: Accurate Single Stage Detector Using Recurrent Rolling Convolution. CVPR 2017.
5	WSSN			76.42 %	84.91 %	71.86 %	0.37 s	GPU @ >3.5 Ghz (Python + C/C++)
Z. Guo, W. Liao, Y. Xiao, P. Veelaert and W. Philips: Weak Segmentation Supervised Deep Neural Networks for Pedestrian Detection. Pattern Recognition 2021.
6	ECP Faster R-CNN			76.25 %	85.96 %	70.55 %	0.25 s	GPU @ 2.5 Ghz (Python)
M. Braun, S. Krebs, F. Flohr and D. Gavrila: The EuroCity Persons Dataset: A Novel Benchmark for Object Detection. CoRR 2018.
7	Aston-EAS			76.07 %	86.71 %	70.02 %	0.24 s	GPU @ 2.5 Ghz (Python + C/C++)
J. Wei, J. He, Y. Zhou, K. Chen, Z. Tang and Z. Xiong: Enhanced Object Detection With Deep Convolutional Neural Networks for Advanced Driving Assistance. IEEE Transactions on Intelligent Transportation Systems 2019.
8	MHN			75.99 %	87.21 %	69.50 %	0.39 s	GPU @ 2.5 Ghz (Python)
J. Cao, Y. Pang, S. Zhao and X. Li: High-Level Semantic Networks for Multi- Scale Object Detection. IEEE Transactions on Circuits and Systems for Video Technology 2019.
9	FFNet		code	75.81 %	87.17 %	69.86 %	1.07 s	GPU @ 1.5 Ghz (Python)
C. Zhao, Y. Qian and M. Yang: Monocular Pedestrian Orientation Estimation Based on Deep 2D-3D Feedforward. Pattern Recognition 2019.
10	SJTU-HW			75.81 %	87.17 %	69.86 %	0.85s	GPU @ 1.5 Ghz (Python + C/C++)
S. Zhang, X. Zhao, L. Fang, F. Haiping and S. Haitao: LED: LOCALIZATION-QUALITY ESTIMATION EMBEDDED DETECTOR. IEEE International Conference on Image Processing 2018. L. Fang, X. Zhao and S. Zhang: Small-objectness sensitive detection based on shifted single shot detector. Multimedia Tools and Applications 2018.
11	MS-CNN		code	74.89 %	85.71 %	68.99 %	0.4 s	GPU @ 2.5 Ghz (C/C++)
Z. Cai, Q. Fan, R. Feris and N. Vasconcelos: A Unified Multi-scale Deep Convolutional Neural Network for Fast Object Detection. ECCV 2016.
12	DD3D		code	73.09 %	85.71 %	68.54 %	n/a s	1 core @ 2.5 Ghz (C/C++)
D. Park, R. Ambrus, V. Guizilini, J. Li and A. Gaidon: Is Pseudo-Lidar needed for Monocular 3D Object detection?. IEEE/CVF International Conference on Computer Vision (ICCV) .
13	F-ConvNet		code	72.91 %	83.63 %	67.18 %	0.47 s	GPU @ 2.5 Ghz (Python + C/C++)
Z. Wang and K. Jia: Frustum ConvNet: Sliding Frustums to Aggregate Local Point-Wise Features for Amodal 3D Object Detection. IROS 2019.
14	GN			72.29 %	82.93 %	65.56 %	1 s	GPU @ 2.5 Ghz (Matlab + C/C++)
S. Jung and K. Hong: Deep network aided by guiding network for pedestrian detection. Pattern Recognition Letters 2017.
15	SubCNN			72.27 %	84.88 %	66.82 %	2 s	GPU @ 3.5 Ghz (Python + C/C++)
Y. Xiang, W. Choi, Y. Lin and S. Savarese: Subcategory-aware Convolutional Neural Networks for Object Proposals and Detection. IEEE Winter Conference on Applications of Computer Vision (WACV) 2017.
16	VMVS			71.82 %	82.80 %	66.85 %	0.25 s	GPU @ 2.5 Ghz (Python)
J. Ku, A. Pon, S. Walsh and S. Waslander: Improving 3D object detection for pedestrians with virtual multi-view synthesis orientation estimation. IROS 2019.
17	EOTL		code	71.45 %	84.74 %	64.58 %	TBD s	1 core @ 2.5 Ghz (Python + C/C++)
R. Yang, Z. Yan, T. Yang, Y. Wang and Y. Ruichek: Efficient Online Transfer Learning for Road Participants Detection in Autonomous Driving. IEEE Sensors Journal 2023.
18	IVA		code	71.37 %	84.61 %	64.90 %	0.4 s	GPU @ 2.5 Ghz (C/C++)
Y. Zhu, J. Wang, C. Zhao, H. Guo and H. Lu: Scale-adaptive Deconvolutional Regression Network for Pedestrian Detection. ACCV 2016. S. Ren, K. He, R. Girshick and J. Sun: Faster R-CNN: Towards real-time object detection with region proposal networks. Advances in neural information processing systems 2015.
19	MM-MRFC			70.76 %	83.79 %	64.81 %	0.05 s	GPU @ 2.5 Ghz (C/C++)
A. Costea, R. Varga and S. Nedevschi: Fast Boosting based Detection using Scale Invariant Multimodal Multiresolution Filtered Features. CVPR 2017.
20	SDP+RPN			70.42 %	82.07 %	65.09 %	0.4 s	GPU @ 2.5 Ghz (Python + C/C++)
F. Yang, W. Choi and Y. Lin: Exploit All the Layers: Fast and Accurate CNN Object Detector with Scale Dependent Pooling and Cascaded Rejection Classifiers. Proceedings of the IEEE International Conference on Computer Vision and Pattern Recognition 2016. S. Ren, K. He, R. Girshick and J. Sun: Faster R-CNN: Towards real-time object detection with region proposal networks. Advances in Neural Information Processing Systems 2015.
21	3DOP		code	69.57 %	83.17 %	63.48 %	3s	GPU @ 2.5 Ghz (Matlab + C/C++)
X. Chen, K. Kundu, Y. Zhu, A. Berneshawi, H. Ma, S. Fidler and R. Urtasun: 3D Object Proposals for Accurate Object Class Detection. NIPS 2015.
22	MonoPSR		code	68.56 %	85.60 %	63.34 %	0.2 s	GPU @ 3.5 Ghz (Python)
J. Ku, A. Pon and S. Waslander: Monocular 3D Object Detection Leveraging Accurate Proposals and Shape Reconstruction. CVPR 2019.
23	DeepStereoOP			68.46 %	83.00 %	63.35 %	3.4 s	GPU @ 3.5 Ghz (Matlab + C/C++)
C. Pham and J. Jeon: Robust Object Proposals Re-ranking for Object Detection in Autonomous Driving Using Convolutional Neural Networks. Signal Processing: Image Communiation 2017.
24	sensekitti		code	68.41 %	82.72 %	62.72 %	4.5 s	GPU @ 2.5 Ghz (Python + C/C++)
B. Yang, J. Yan, Z. Lei and S. Li: Craft Objects from Images. CVPR 2016.
25	MonoAFKD			67.83 %	82.92 %	60.90 %	0.03 s	1 core @ 2.5 Ghz (C/C++)

26	MonoLSS			67.78 %	82.88 %	60.87 %	0.04 s	1 core @ 2.5 Ghz (Python)
Z. Li, J. Jia and Y. Shi: MonoLSS: Learnable Sample Selection For Monocular 3D Detection. International Conference on 3D Vision 2024.
27	Frustum-PointPillars		code	67.51 %	76.80 %	63.81 %	0.06 s	4 cores @ 3.0 Ghz (Python)
A. Paigwar, D. Sierra-Gonzalez, \. Erkent and C. Laugier: Frustum-PointPillars: A Multi-Stage Approach for 3D Object Detection using RGB Camera and LiDAR. International Conference on Computer Vision, ICCV, Workshop on Autonomous Vehicle Vision 2021.
28	FII-CenterNet		code	67.31 %	81.32 %	61.29 %	0.09 s	GPU @ 2.5 Ghz (Python)
S. Fan, F. Zhu, S. Chen, H. Zhang, B. Tian, Y. Lv and F. Wang: FII-CenterNet: An Anchor-Free Detector With Foreground Attention for Traffic Object Detection. IEEE Transactions on Vehicular Technology 2021.
29	Mono3D		code	67.29 %	80.30 %	62.23 %	4.2 s	GPU @ 2.5 Ghz (Matlab + C/C++)
X. Chen, K. Kundu, Z. Zhang, H. Ma, S. Fidler and R. Urtasun: Monocular 3D Object Detection for Autonomous Driving. CVPR 2016.
30	Faster R-CNN		code	66.24 %	79.97 %	61.09 %	2 s	GPU @ 3.5 Ghz (Python + C/C++)
S. Ren, K. He, R. Girshick and J. Sun: Faster R-CNN: Towards Real- Time Object Detection with Region Proposal Networks. NIPS 2015.
31	VPFNet		code	65.68 %	75.03 %	61.95 %	0.2 s	1 core @ 2.5 Ghz (C/C++)
C. Wang, H. Chen and L. Fu: VPFNet: Voxel-Pixel Fusion Network for Multi-class 3D Object Detection. 2021. C. Wang, H. Chen, Y. Chen, P. Hsiao and L. Fu: VoPiFNet: Voxel-Pixel Fusion Network for Multi-Class 3D Object Detection. IEEE Transactions on Intelligent Transportation Systems 2024.
32	UPIDet		code	65.50 %	75.07 %	63.09 %	0.11 s	1 core @ 2.5 Ghz (C/C++)
Y. Zhang, Q. Zhang, J. Hou, Y. Yuan and G. Xing: Unleash the Potential of Image Branch for Cross-modal 3D Object Detection. Thirty-seventh Conference on Neural Information Processing Systems 2023.
33	EQ-PVRCNN		code	65.01 %	77.19 %	61.95 %	0.2 s	GPU @ 2.5 Ghz (Python + C/C++)
Z. Yang, L. Jiang, Y. Sun, B. Schiele and J. Jia: A Unified Query-based Paradigm for Point Cloud Understanding. Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition 2022.
34	CasA++		code	64.94 %	74.41 %	62.35 %	0.1 s	1 core @ 2.5 Ghz (C/C++)
H. Wu, J. Deng, C. Wen, X. Li and C. Wang: CasA: A Cascade Attention Network for 3D Object Detection from LiDAR point clouds. IEEE Transactions on Geoscience and Remote Sensing 2022.
35	TED		code	64.74 %	74.26 %	62.08 %	0.1 s	1 core @ 2.5 Ghz (C/C++)
H. Wu, C. Wen, W. Li, R. Yang and C. Wang: Transformation-Equivariant 3D Object Detection for Autonomous Driving. AAAI 2023.
36	LoGoNet		code	64.55 %	72.47 %	62.24 %	0.1 s	1 core @ 2.5 Ghz (C/C++)
X. Li, T. Ma, Y. Hou, B. Shi, Y. Yang, Y. Liu, X. Wu, Q. Chen, Y. Li, Y. Qiao and others: LoGoNet: Towards Accurate 3D Object Detection with Local-to-Global Cross-Modal Fusion. CVPR 2023.
37	SDP+CRC (ft)			64.36 %	79.22 %	59.16 %	0.6 s	GPU @ 2.5 Ghz (C/C++)
F. Yang, W. Choi and Y. Lin: Exploit All the Layers: Fast and Accurate CNN Object Detector with Scale Dependent Pooling and Cascaded Rejection Classifiers. Proceedings of the IEEE International Conference on Computer Vision and Pattern Recognition 2016.
38	LCANet			64.13 %	73.98 %	60.14 %	1 s	1 core @ 2.5 Ghz (C/C++)

39	Pose-RCNN			63.54 %	80.07 %	57.02 %	2 s	>8 cores @ 2.5 Ghz (Python)
M. Braun, Q. Rao, Y. Wang and F. Flohr: Pose-RCNN: Joint object detection and pose estimation using 3D object proposals. Intelligent Transportation Systems (ITSC), 2016 IEEE 19th International Conference on 2016.
40	USVLab BSAODet		code	63.21 %	72.86 %	59.48 %	0.04 s	1 core @ 2.5 Ghz (C/C++)
W. Xiao, Y. Peng, C. Liu, J. Gao, Y. Wu and X. Li: Balanced Sample Assignment and Objective for Single-Model Multi-Class 3D Object Detection. IEEE Transactions on Circuits and Systems for Video Technology 2023.
41	MLF-DET			63.09 %	70.25 %	59.23 %	0.09 s	1 core @ 2.5 Ghz (C/C++)
Z. Lin, Y. Shen, S. Zhou, S. Chen and N. Zheng: MLF-DET: Multi-Level Fusion for Cross- Modal 3D Object Detection. International Conference on Artificial Neural Networks 2023.
42	FIRM-Net_SCF+			62.91 %	72.98 %	60.31 %	0.07 s	1 core @ 2.5 Ghz (Python)

43	CFM			62.84 %	74.76 %	56.06 %	<2 s	GPU @ 2.5 Ghz (Matlab + C/C++)
Q. Hu, P. Wang, C. Shen, A. Hengel and F. Porikli: Pushing the Limits of Deep CNNs for Pedestrian Detection. IEEE Transactions on Circuits and Systems for Video Technology 2017.
44	CasA		code	62.73 %	72.65 %	60.12 %	0.1 s	1 core @ 2.5 Ghz (C/C++)
H. Wu, J. Deng, C. Wen, X. Li and C. Wang: CasA: A Cascade Attention Network for 3D Object Detection from LiDAR point clouds. IEEE Transactions on Geoscience and Remote Sensing 2022.
45	FIRM-Net-SCF			62.63 %	72.78 %	59.99 %	0.07 s	1 core @ 2.5 Ghz (Python)

46	Fast-CLOCs			62.57 %	76.22 %	60.13 %	0.1 s	GPU @ 2.5 Ghz (Python)
S. Pang, D. Morris and H. Radha: Fast-CLOCs: Fast Camera-LiDAR Object Candidates Fusion for 3D Object Detection. Proceedings of the IEEE/CVF Winter Conference on Applications of Computer Vision (WACV) 2022.
47	PiFeNet		code	62.35 %	72.74 %	59.29 %	0.03 s	1 core @ 2.5 Ghz (C/C++)
D. Le, H. Shi, H. Rezatofighi and J. Cai: Accurate and Real-time 3D Pedestrian Detection Using an Efficient Attentive Pillar Network. IEEE Robotics and Automation Letters 2022.
48	HotSpotNet			62.31 %	71.43 %	59.24 %	0.04 s	1 core @ 2.5 Ghz (Python + C/C++)
Q. Chen, L. Sun, Z. Wang, K. Jia and A. Yuille: object as hotspots. Proceedings of the European Conference on Computer Vision (ECCV) 2020.
49	BVIFusion+			62.06 %	71.81 %	58.21 %	0.09 s	1 core @ 2.5 Ghz (Python)

50	P2V-RCNN			61.83 %	71.76 %	59.29 %	0.1 s	2 cores @ 2.5 Ghz (Python)
J. Li, S. Luo, Z. Zhu, H. Dai, A. Krylov, Y. Ding and L. Shao: P2V-RCNN: Point to Voxel Feature Learning for 3D Object Detection from Point Clouds. IEEE Access 2021.
51	MonoPair			61.57 %	78.81 %	56.51 %	0.06 s	GPU @ 2.5 Ghz (Python + C/C++)
Y. Chen, L. Tai, K. Sun and M. Li: MonoPair: Monocular 3D Object Detection Using Pairwise Spatial Relationships. The IEEE Conference on Computer Vision and Pattern Recognition (CVPR) 2020.
52	LumiNet		code	61.38 %	72.01 %	58.94 %	0.1 s	1 core @ 2.5 Ghz (Python)

53	monodle		code	61.29 %	78.66 %	56.18 %	0.04 s	GPU @ 2.5 Ghz (Python)
X. Ma, Y. Zhang, D. Xu, D. Zhou, S. Yi, H. Li and W. Ouyang: Delving into Localization Errors for Monocular 3D Object Detection. CVPR 2021 .
54	OGMMDet		code	61.26 %	72.41 %	58.79 %	0.01 s	1 core @ 2.5 Ghz (C/C++)

55	ANM		code	61.26 %	72.41 %	58.79 %	ANM	ANM

56	RPN+BF		code	61.22 %	77.06 %	55.22 %	0.6 s	GPU @ 2.5 Ghz (Matlab + C/C++)
L. Zhang, L. Lin, X. Liang and K. He: Is Faster R-CNN Doing Well for Pedestrian Detection?. ECCV 2016.
57	AFFN-G			60.99 %	69.06 %	58.89 %	0.5 s	1 core @ 2.5 Ghz (C/C++)

58	GEFPN			60.99 %	69.06 %	58.89 %	0.5 s	1 core @ 2.5 Ghz (C/C++)

59	GeVo			60.99 %	69.06 %	58.89 %	0.05 s	1 core @ 2.5 Ghz (C/C++)

60	3ONet			60.89 %	72.45 %	56.65 %	0.1 s	1 core @ 2.5 Ghz (C/C++)
H. Hoang and M. Yoo: 3ONet: 3-D Detector for Occluded Object Under Obstructed Conditions. IEEE Sensors Journal 2023.
61	Regionlets			60.83 %	73.79 %	54.72 %	1 s	>8 cores @ 2.5 Ghz (C/C++)
X. Wang, M. Yang, S. Zhu and Y. Lin: Regionlets for Generic Object Detection. T-PAMI 2015. W. Zou, X. Wang, M. Sun and Y. Lin: Generic Object Detection with Dense Neural Patterns and Regionlets. British Machine Vision Conference 2014. C. Long, X. Wang, G. Hua, M. Yang and Y. Lin: Accurate Object Detection with Location Relaxation and Regionlets Relocalization. Asian Conference on Computer Vision 2014.
62	PillarHist			60.78 %	71.70 %	57.66 %	0.01 s	1 core @ 2.5 Ghz (C/C++)

63	3DSSD		code	60.51 %	72.33 %	56.28 %	0.04 s	GPU @ 2.5 Ghz (Python + C/C++)
Z. Yang, Y. Sun, S. Liu and J. Jia: 3DSSD: Point-based 3D Single Stage Object Detector. CVPR 2020.
64	ACFNet			60.12 %	71.42 %	55.96 %	0.11 s	1 core @ 2.5 Ghz (C/C++)
Y. Tian, X. Zhang, X. Wang, J. Xu, J. Wang, R. Ai, W. Gu and W. Ding: ACF-Net: Asymmetric Cascade Fusion for 3D Detection With LiDAR Point Clouds and Images. IEEE Transactions on Intelligent Vehicles 2023.
65	vsis-PHNet			60.10 %	71.15 %	57.66 %	0.03 s	1 core @ 2.5 Ghz (C/C++)

66	PHNetp			60.10 %	71.15 %	57.66 %	0.03 s	1 core @ 2.5 Ghz (C/C++)

67	dsvd+vx			59.75 %	69.79 %	57.34 %	0.1 s	1 core @ 2.5 Ghz (C/C++)

68	DPPFA-Net			59.52 %	67.68 %	56.87 %	0.1 s	1 core @ 2.5 Ghz (Python)
J. Wang, X. Kong, H. Nishikawa, Q. Lian and H. Tomiyama: Dynamic Point-Pixel Feature Alignment for Multi-modal 3D Object Detection. IEEE Internet of Things Journal 2023.
69	ACDet		code	59.51 %	71.27 %	57.03 %	0.05 s	1 core @ 2.5 Ghz (C/C++)
J. Xu, G. Wang, X. Zhang and G. Wan: ACDet: Attentive Cross-view Fusion for LiDAR-based 3D Object Detection. 3DV 2022.
70	LPFusion_three_class			59.48 %	69.28 %	55.55 %	0.1 s	1 core @ 2.5 Ghz (C/C++)

71	SCNet3D			59.47 %	69.32 %	56.96 %	0.08 s	1 core @ 2.5 Ghz (Python)

72	QD-3DT		code	59.26 %	78.41 %	54.37 %	0.03 s	GPU @ 2.5 Ghz (Python)
H. Hu, Y. Yang, T. Fischer, F. Yu, T. Darrell and M. Sun: Monocular Quasi-Dense 3D Object Tracking. ArXiv:2103.07351 2021.
73	AM			59.09 %	77.32 %	54.25 %	0.04 s	1 core @ 2.5 Ghz (C/C++)

74	TANet		code	59.07 %	69.90 %	56.44 %	0.035s	GPU @ 2.5 Ghz (Python + C/C++)
Z. Liu, X. Zhao, T. Huang, R. Hu, Y. Zhou and X. Bai: TANet: Robust 3D Object Detection from Point Clouds with Triple Attention. AAAI 2020.
75	MonoUNI		code	58.97 %	76.17 %	53.99 %	0.04 s	1 core @ 2.5 Ghz (Python)
J. Jia, Z. Li and Y. Shi: MonoUNI: A Unified Vehicle and Infrastructure-side Monocular 3D Object Detection Network with Sufficient Depth Clues. Thirty-seventh Conference on Neural Information Processing Systems 2023.
76	SFA_IGCL_Focalsconv*		code	58.94 %	66.88 %	56.80 %	0.2 s	1 core @ 2.5 Ghz (C/C++)

77	SVFMamba		code	58.83 %	69.22 %	55.36 %	N/A s	1 core @ 2.5 Ghz (C/C++)

78	DSA-PV-RCNN		code	58.81 %	66.93 %	56.57 %	0.08 s	1 core @ 2.5 Ghz (Python + C/C++)
P. Bhattacharyya, C. Huang and K. Czarnecki: SA-Det3D: Self-Attention Based Context-Aware 3D Object Detection. 2021.
79	SRDL			58.70 %	68.45 %	56.23 %	0.05 s	1 core @ 2.5 Ghz (Python + C/C++)
ERROR: Wrong syntax in BIBTEX file.
80	FocalsConv*			58.61 %	66.08 %	55.20 %	0.2 s	1 core @ 2.5 Ghz (C/C++)

81	MMLab PV-RCNN		code	58.37 %	68.88 %	55.38 %	0.08 s	1 core @ 2.5 Ghz (Python + C/C++)
S. Shi, C. Guo, L. Jiang, Z. Wang, J. Shi, X. Wang and H. Li: PV-RCNN: Point-Voxel Feature Set Abstraction for 3D Object Detection. CVPR 2020.
82	PASS-PV-RCNN-Plus			58.31 %	67.45 %	55.92 %	1 s	1 core @ 2.5 Ghz (Python)
Anonymous: Leveraging Anchor-based LiDAR 3D Object Detection via Point Assisted Sample Selection. will submit to computer vision conference/journal 2024.
83	Point-GNN		code	58.20 %	71.59 %	54.06 %	0.6 s	GPU @ 2.5 Ghz (Python)
W. Shi and R. Rajkumar: Point-GNN: Graph Neural Network for 3D Object Detection in a Point Cloud. CVPR 2020.
84	MPC3DNet			58.20 %	66.33 %	55.97 %	0.05 s	GPU @ 1.5 Ghz (Python)

85	test			58.19 %	66.17 %	56.16 %	0.04 s	GPU @ 1.5 Ghz (Python + C/C++)

86	DeepParts			58.15 %	71.47 %	51.92 %	~1 s	GPU @ 2.5 Ghz (Matlab)
Y. Tian, P. Luo, X. Wang and X. Tang: Deep Learning Strong Parts for Pedestrian Detection. ICCV 2015.
87	CompACT-Deep			58.14 %	70.93 %	52.29 %	1 s	1 core @ 2.5 Ghz (Matlab + C/C++)
Z. Cai, M. Saberian and N. Vasconcelos: Learning Complexity-Aware Cascades for Deep Pedestrian Detection. ICCV 2015.
88	EPNet++			58.10 %	68.58 %	55.58 %	0.1 s	GPU @ 2.5 Ghz (Python)
Z. Liu, T. Huang, B. Li, X. Chen, X. Wang and X. Bai: EPNet++: Cascade Bi-Directional Fusion for Multi-Modal 3D Object Detection. IEEE Transactions on Pattern Analysis and Machine Intelligence 2022.
89	DSGN++		code	58.09 %	69.70 %	54.45 %	0.2 s	GeForce RTX 2080Ti
Y. Chen, S. Huang, S. Liu, B. Yu and J. Jia: DSGN++: Exploiting Visual-Spatial Relation for Stereo-Based 3D Detectors. IEEE Transactions on Pattern Analysis and Machine Intelligence 2022.
90	MMLab-PartA^2		code	57.96 %	68.78 %	54.01 %	0.08 s	GPU @ 2.5 Ghz (Python + C/C++)
S. Shi, Z. Wang, J. Shi, X. Wang and H. Li: From Points to Parts: 3D Object Detection from Point Cloud with Part-aware and Part-aggregation Network. IEEE Transactions on Pattern Analysis and Machine Intelligence 2020.
91	SVGA-Net			57.92 %	67.81 %	55.25 %	0.03s	1 core @ 2.5 Ghz (Python + C/C++)
Q. He, Z. Wang, H. Zeng, Y. Zeng and Y. Liu: SVGA-Net: Sparse Voxel-Graph Attention Network for 3D Object Detection from Point Clouds. AAAI 2022.
92	AVOD-FPN		code	57.87 %	67.95 %	55.23 %	0.1 s	Titan X (Pascal)
J. Ku, M. Mozifian, J. Lee, A. Harakeh and S. Waslander: Joint 3D Proposal Generation and Object Detection from View Aggregation. IROS 2018.
93	HA-PillarNet			57.87 %	65.21 %	55.84 %	0.05 s	1 core @ 2.5 Ghz (Python)

94	Voxel RCNN-Focal*		code	57.74 %	65.53 %	55.67 %	0.2 s	1 core @ 2.5 Ghz (Python)

95	DFAF3D			57.65 %	67.45 %	53.89 %	0.05 s	1 core @ 2.5 Ghz (Python)
Q. Tang, X. Bai, J. Guo, B. Pan and W. Jiang: DFAF3D: A dual-feature-aware anchor-free single-stage 3D detector for point clouds. Image and Vision Computing 2023.
96	AFFN			57.59 %	66.29 %	55.14 %	0.5 s	1 core @ 2.5 Ghz (C/C++)

97	CGML			57.49 %	64.84 %	55.41 %	0.33 s	1 core @ 2.5 Ghz (C/C++)

98	Faraway-Frustum		code	57.35 %	67.88 %	54.42 %	0.1 s	GPU @ 2.5 Ghz (Python)
H. Zhang, D. Yang, E. Yurtsever, K. Redmill and U. Ozguner: Faraway-frustum: Dealing with lidar sparsity for 3D object detection using fusion. 2021 IEEE International Intelligent Transportation Systems Conference (ITSC) 2021.
99	PDV		code	57.34 %	65.94 %	54.21 %	0.1 s	1 core @ 2.5 Ghz (C/C++)
J. Hu, T. Kuai and S. Waslander: Point Density-Aware Voxels for LiDAR 3D Object Detection. CVPR 2022.
100	SIF			57.32 %	67.78 %	54.86 %	0.1 s	1 core @ 2.5 Ghz (C/C++)
P. An: SIF. Submitted to CVIU 2021.
101	PG-RCNN		code	57.31 %	67.77 %	54.83 %	0.06 s	GPU @ 1.5 Ghz (Python)
I. Koo, I. Lee, S. Kim, H. Kim, W. Jeon and C. Kim: PG-RCNN: Semantic Surface Point Generation for 3D Object Detection. 2023.
102	FromVoxelToPoint		code	57.26 %	68.26 %	54.74 %	0.1 s	1 core @ 2.5 Ghz (C/C++)
J. Li, H. Dai, L. Shao and Y. Ding: From Voxel to Point: IoU-guided 3D Object Detection for Point Cloud with Voxel-to- Point Decoder. MM '21: The 29th ACM International Conference on Multimedia (ACM MM) 2021.
103	MLFusion-VS			57.22 %	66.13 %	55.00 %	0.06 s	1 core @ 2.5 Ghz (C/C++)

104	SemanticVoxels			57.22 %	67.62 %	54.90 %	0.04 s	GPU @ 2.5 Ghz (Python + C/C++)
J. Fei, W. Chen, P. Heidenreich, S. Wirges and C. Stiller: SemanticVoxels: Sequential Fusion for 3D Pedestrian Detection using LiDAR Point Cloud and Semantic Segmentation. MFI 2020.
105	IA-SSD (single)		code	56.87 %	66.69 %	54.68 %	0.013 s	1 core @ 2.5 Ghz (C/C++)
Y. Zhang, Q. Hu, G. Xu, Y. Ma, J. Wan and Y. Guo: Not All Points Are Equal: Learning Highly Efficient Point-based Detectors for 3D LiDAR Point Clouds. CVPR 2022.
106	SFA-GCL(80, k=4)		code	56.83 %	69.60 %	54.42 %	0.04 s	1 core @ 2.5 Ghz (C/C++)

107	CAT-Det			56.75 %	67.15 %	53.44 %	0.3 s	GPU @ 2.5 Ghz (Python + C/C++)
Y. Zhang, J. Chen and D. Huang: CAT-Det: Contrastively Augmented Transformer for Multi-modal 3D Object Detection. CVPR 2022.
108	SFA-GCL_dataaug		code	56.73 %	69.55 %	54.32 %	0.04 s	1 core @ 2.5 Ghz (C/C++)

109	SFA-GCL(80)		code	56.69 %	67.64 %	52.51 %	0.04 s	1 core @ 2.5 Ghz (C/C++)

110	FRCNN+Or		code	56.68 %	71.64 %	51.53 %	0.09 s	Titan Xp GPU
C. Guindel, D. Martin and J. Armingol: Fast Joint Object Detection and Viewpoint Estimation for Traffic Scene Understanding. IEEE Intelligent Transportation Systems Magazine 2018. C. Guindel, D. Martin and J. Armingol: Joint Object Detection and Viewpoint Estimation using CNN features. IEEE International Conference on Vehicular Electronics and Safety (ICVES) 2017.
111	SFA-GCL		code	56.62 %	69.30 %	54.20 %	0.04 s	1 core @ 2.5 Ghz (C/C++)

112	FilteredICF			56.53 %	69.79 %	50.32 %	~ 2 s	>8 cores @ 2.5 Ghz (Matlab + C/C++)
S. Zhang, R. Benenson and B. Schiele: Filtered Channel Features for Pedestrian Detection. CVPR 2015.
113	R2Pfusion-Det			56.52 %	67.09 %	54.46 %	0.3 s	1 core @ 2.5 Ghz (C/C++)

114	R50_SACINet			56.47 %	67.56 %	52.71 %	0.06 s	1 core @ 2.5 Ghz (C/C++)

115	HMNet			56.46 %	67.93 %	53.96 %	0.1 s	1 core @ 2.5 Ghz (C/C++)

116	SFA-GCL(baseline)		code	56.42 %	69.02 %	54.05 %	0.04 s	1 core @ 2.5 Ghz (C/C++)

117	ARPNET			56.42 %	69.08 %	52.69 %	0.08 s	GPU @ 2.5 Ghz (Python + C/C++)
Y. Ye, C. Zhang and X. Hao: ARPNET: attention region proposal network for 3D object detection. Science China Information Sciences 2019.
118	MonoRUn		code	56.40 %	73.05 %	51.40 %	0.07 s	GPU @ 2.5 Ghz (Python + C/C++)
H. Chen, Y. Huang, W. Tian, Z. Gao and L. Xiong: MonoRUn: Monocular 3D Object Detection by Reconstruction and Uncertainty Propagation. IEEE Conference on Computer Vision and Pattern Recognition (CVPR) 2021.
119	MV-RGBD-RF			56.18 %	72.99 %	49.72 %	4 s	4 cores @ 2.5 Ghz (C/C++)
A. Gonzalez, D. Vazquez, A. Lopez and J. Amores: On-Board Object Detection: Multicue, Multimodal, and Multiview Random Forest of Local Experts.. IEEE Trans. on Cybernetics 2016. A. Gonzalez, G. Villalonga, J. Xu, D. Vazquez, J. Amores and A. Lopez: Multiview Random Forest of Local Experts Combining RGB and LIDAR data for Pedestrian Detection. IEEE Intelligent Vehicles Symposium (IV) 2015.
120	CG-SSD			55.97 %	65.50 %	53.64 %	0.01 s	1 core @ 2.5 Ghz (C/C++)

121	HMFI		code	55.96 %	66.20 %	53.24 %	0.1 s	1 core @ 2.5 Ghz (C/C++)
X. Li, B. Shi, Y. Hou, X. Wu, T. Ma, Y. Li and L. He: Homogeneous Multi-modal Feature Fusion and Interaction for 3D Object Detection. ECCV 2022.
122	MGAF-3DSSD		code	55.80 %	66.31 %	52.02 %	0.1 s	1 core @ 2.5 Ghz (C/C++)
J. Li, H. Dai, L. Shao and Y. Ding: Anchor-free 3D Single Stage Detector with Mask-Guided Attention for Point Cloud. MM '21: The 29th ACM International Conference on Multimedia (ACM MM) 2021.
123	GUPNet		code	55.65 %	74.95 %	48.44 %	NA s	1 core @ 2.5 Ghz (Python + C/C++)
Y. Lu, X. Ma, L. Yang, T. Zhang, Y. Liu, Q. Chu, J. Yan and W. Ouyang: Geometry Uncertainty Projection Network for Monocular 3D Object Detection. arXiv preprint arXiv:2107.13774 2021.
124	MLOD		code	55.62 %	68.42 %	51.45 %	0.12 s	GPU @ 1.5 Ghz (Python)
J. Deng and K. Czarnecki: MLOD: A multi-view 3D object detection based on robust feature fusion method. arXiv preprint arXiv:1909.04163 2019.
125	AFFN-Ga			55.60 %	64.75 %	53.23 %	0.5 s	1 core @ 2.5 Ghz (C/C++)

126	L_SACINet			55.37 %	65.81 %	51.63 %	0.04 s	1 core @ 2.5 Ghz (C/C++)

127	SAKD-MR-Res18			55.20 %	74.40 %	48.08 %	0.03 s	1 core @ 2.5 Ghz (C/C++)

128	MonoMH		code	55.19 %	73.25 %	50.19 %	0.04 s	1 core @ 2.5 Ghz (Python)

129	DEVIANT		code	55.16 %	74.27 %	50.21 %	0.04 s	1 GPU (Python)
A. Kumar, G. Brazil, E. Corona, A. Parchami and X. Liu: DEVIANT: Depth EquiVarIAnt NeTwork for Monocular 3D Object Detection. European Conference on Computer Vision (ECCV) 2022.
130	PointPillars		code	55.10 %	65.29 %	52.39 %	16 ms	1080ti GPU and Intel i7 CPU
A. Lang, S. Vora, H. Caesar, L. Zhou, J. Yang and O. Beijbom: PointPillars: Fast Encoders for Object Detection from Point Clouds. CVPR 2019.
131	StereoDistill			55.09 %	69.00 %	50.95 %	0.4 s	1 core @ 2.5 Ghz (Python)
Z. Liu, X. Ye, X. Tan, D. Errui, Y. Zhou and X. Bai: StereoDistill: Pick the Cream from LiDAR for Distilling Stereo-based 3D Object Detection. Proceedings of the AAAI Conference on Artificial Intelligence 2023.
132	STD		code	55.04 %	68.33 %	50.85 %	0.08 s	GPU @ 2.5 Ghz (Python + C/C++)
Z. Yang, Y. Sun, S. Liu, X. Shen and J. Jia: STD: Sparse-to-Dense 3D Object Detector for Point Cloud. ICCV 2019.
133	OPA-3D		code	54.92 %	73.93 %	47.87 %	0.04 s	1 core @ 3.5 Ghz (Python)
Y. Su, Y. Di, G. Zhai, F. Manhardt, J. Rambach, B. Busam, D. Stricker and F. Tombari: OPA-3D: Occlusion-Aware Pixel-Wise Aggregation for Monocular 3D Object Detection. IEEE Robotics and Automation Letters 2023.
134	Vote3Deep			54.80 %	67.99 %	51.17 %	1.5 s	4 cores @ 2.5 Ghz (C/C++)
M. Engelcke, D. Rao, D. Zeng Wang, C. Hay Tong and I. Posner: Vote3Deep: Fast Object Detection in 3D Point Clouds Using Efficient Convolutional Neural Networks. ArXiv e-prints 2016.
135	M3DeTR		code	54.78 %	63.15 %	52.49 %	n/a s	GPU @ 1.0 Ghz (Python)
T. Guan, J. Wang, S. Lan, R. Chandra, Z. Wu, L. Davis and D. Manocha: M3DeTR: Multi-representation, Multi- scale, Mutual-relation 3D Object Detection with Transformers. 2021.
136	MFB3D			54.75 %	62.76 %	51.89 %	0.14 s	1 core @ 2.5 Ghz (Python)

137	CAIA_PRO		code	54.70 %	64.33 %	52.36 %	0.01 s	1 core @ 2.5 Ghz (C/C++)

138	L-AUG			54.61 %	65.71 %	51.67 %	0.1 s	1 core @ 2.5 Ghz (C/C++)
T. Cortinhal, I. Gouigah and E. Aksoy: Semantics-aware LiDAR-Only Pseudo Point Cloud Generation for 3D Object Detection. 2023.
139	SecAtten			54.61 %	65.63 %	51.02 %	0.1 s	1 core @ 2.5 Ghz (C/C++)

140	MonoCoP			54.41 %	70.75 %	49.66 %	0.01 s	1 core @ 2.5 Ghz (C/C++)

141	AEPF			54.41 %	64.95 %	51.22 %	0.05 s	GPU @ 2.5 Ghz (Python)

142	SFA-GCL		code	54.27 %	66.81 %	50.13 %	0.04 s	1 core @ 2.5 Ghz (C/C++)

143	epBRM		code	54.13 %	62.90 %	51.95 %	0.10 s	1 core @ 2.5 Ghz (C/C++)
K. Shin: Improving a Quality of 3D Object Detection by Spatial Transformation Mechanism. arXiv preprint arXiv:1910.04853 2019.
144	DVFENet			54.13 %	63.54 %	51.79 %	0.05 s	1 core @ 2.5 Ghz (Python + C/C++)
Y. He, G. Xia, Y. Luo, L. Su, Z. Zhang, W. Li and P. Wang: DVFENet: Dual-branch Voxel Feature Extraction Network for 3D Object Detection. Neurocomputing 2021.
145	SA V1			54.09 %	64.15 %	51.64 %	0.5 s	GPU @ 2.5 Ghz (Python)

146	XView			53.83 %	62.27 %	51.61 %	0.1 s	1 core @ 2.5 Ghz (C/C++)
L. Xie, G. Xu, D. Cai and X. He: X-view: Non-egocentric Multi-View 3D Object Detector. 2021.
147	monospb			53.79 %	72.50 %	48.89 %	0.01 s	1 core @ 2.5 Ghz (Python)

148	PointPainting			53.76 %	61.86 %	50.61 %	0.4 s	GPU @ 2.5 Ghz (Python + C/C++)
S. Vora, A. Lang, B. Helou and O. Beijbom: PointPainting: Sequential Fusion for 3D Object Detection. CVPR 2020.
149	PDV2			53.54 %	65.59 %	47.65 %	3.7 s	1 core @ 3.0 Ghz Matlab (C/C++)
J. Shen, X. Zuo, J. Li, W. Yang and H. Ling: A novel pixel neighborhood differential statistic feature for pedestrian and face detection . Pattern Recognition 2017.
150	Mix-Teaching		code	53.52 %	67.34 %	47.45 %	30 s	1 core @ 2.5 Ghz (C/C++)
L. Yang, X. Zhang, L. Wang, M. Zhu, C. Zhang and J. Li: Mix-Teaching: A Simple, Unified and Effective Semi-Supervised Learning Framework for Monocular 3D Object Detection. ArXiv 2022.
151	Cube R-CNN		code	53.27 %	64.96 %	47.84 %	0.05 s	GPU @ 2.5 Ghz (Python)
G. Brazil, A. Kumar, J. Straub, N. Ravi, J. Johnson and G. Gkioxari: Omni3D: A Large Benchmark and Model for 3D Object Detection in the Wild. CVPR 2023.
152	TAFT			53.15 %	67.62 %	47.08 %	0.2 s	1 core @ 2.5 Ghz (C/C++)
J. Shen, X. Zuo, W. Yang, D. Prokhorov, X. Mei and H. Ling: Differential Features for Pedestrian Detection: A Taylor Series Perspective. IEEE Transactions on Intelligent Transportation Systems 2018.
153	MSMA V1			53.10 %	62.22 %	50.28 %	0.5 s	GPU @ 2.5 Ghz (Python)

154	Disp R-CNN		code	52.98 %	71.79 %	48.20 %	0.387 s	GPU @ 2.5 Ghz (Python + C/C++)
J. Sun, L. Chen, Y. Xie, S. Zhang, Q. Jiang, X. Zhou and H. Bao: Disp R-CNN: Stereo 3D Object Detection via Shape Prior Guided Instance Disparity Estimation. CVPR 2020.
155	Disp R-CNN (velo)		code	52.90 %	71.63 %	48.15 %	0.387 s	GPU @ 2.5 Ghz (Python + C/C++)
J. Sun, L. Chen, Y. Xie, S. Zhang, Q. Jiang, X. Zhou and H. Bao: Disp R-CNN: Stereo 3D Object Detection via Shape Prior Guided Instance Disparity Estimation. CVPR 2020.
156	pAUCEnsT			52.88 %	65.84 %	46.97 %	60 s	1 core @ 2.5 Ghz (Matlab + C/C++)
S. Paisitkriangkrai, C. Shen and A. Hengel: Pedestrian Detection with Spatially Pooled Features and Structured Ensemble Learning. arXiv 2014.
157	SparVox3D			52.84 %	69.33 %	48.49 %	0.05 s	GPU @ 2.0 Ghz (Python)
E. Balatkan and F. Kıraç: Improving Regression Performance on Monocular 3D Object Detection Using Bin-Mixing and Sparse Voxel Data. 2021 6th International Conference on Computer Science and Engineering (UBMK) 2021.
158	GSG-FPS		code	52.66 %	61.73 %	50.50 %	0.01 s	1 core @ 2.5 Ghz (C/C++)

159	PFF3D		code	52.53 %	62.12 %	50.27 %	0.05 s	GPU @ 3.0 Ghz (Python + C/C++)
L. Wen and K. Jo: Fast and Accurate 3D Object Detection for Lidar-Camera-Based Autonomous Vehicles Using One Shared Voxel-Based Backbone. IEEE Access 2021.
160	IA-SSD (multi)		code	52.45 %	65.07 %	50.20 %	0.014 s	1 core @ 2.5 Ghz (C/C++)
Y. Zhang, Q. Hu, G. Xu, Y. Ma, J. Wan and Y. Guo: Not All Points Are Equal: Learning Highly Efficient Point-based Detectors for 3D LiDAR Point Clouds. CVPR 2022.
161	S-AT GCN			52.30 %	62.01 %	50.10 %	0.02 s	GPU @ 2.0 Ghz (Python)
L. Wang, C. Wang, X. Zhang, T. Lan and J. Li: S-AT GCN: Spatial-Attention Graph Convolution Network based Feature Enhancement for 3D Object Detection. CoRR 2021.
162	MMLAB LIGA-Stereo		code	52.18 %	65.59 %	49.29 %	0.4 s	1 core @ 2.5 Ghz (Python + C/C++)
X. Guo, S. Shi, X. Wang and H. Li: LIGA-Stereo: Learning LiDAR Geometry Aware Representations for Stereo-based 3D Detector. Proceedings of the IEEE/CVF International Conference on Computer Vision (ICCV) 2021.
163	HINTED		code	51.95 %	66.52 %	47.83 %	0.04 s	1 core @ 2.5 Ghz (C/C++)
Q. Xia, W. Ye, H. Wu, S. Zhao, L. Xing, X. Huang, J. Deng, X. Li, C. Wen and C. Wang: HINTED: Hard Instance Enhanced Detector with Mixed-Density Feature Fusion for Sparsely- Supervised 3D Object Detection. Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition 2024.
164	PCNet3D++			51.69 %	62.37 %	49.14 %	0.5 s	GPU @ 3.5 Ghz (Python)

165	PCNet3D			51.69 %	61.97 %	49.20 %	0.05 s	GPU @ 3.5 Ghz (Python)

166	Plane-Constraints		code	51.57 %	64.64 %	46.98 %	0.05 s	4 cores @ 3.0 Ghz (Python)
H. Yao, J. Chen, Z. Wang, X. Wang, X. Chai, Y. Qiu and P. Han: Vertex points are not enough: Monocular 3D object detection via intra-and inter-plane constraints. Neural Networks 2023.
167	Anonymous			51.42 %	62.54 %	48.94 %	0.02 s	GPU @ 2.5 Ghz (Python)

168	Test_dif		code	51.35 %	60.86 %	49.27 %	0.01 s	1 core @ 2.5 Ghz (C/C++)

169	Shift R-CNN (mono)		code	51.30 %	70.86 %	46.37 %	0.25 s	GPU @ 1.5 Ghz (Python)
A. Naiden, V. Paunescu, G. Kim, B. Jeon and M. Leordeanu: Shift R-CNN: Deep Monocular 3D Object Detection With Closed-form Geometric Constraints. ICIP 2019.
170	SeSame-point		code	51.27 %	60.29 %	49.06 %	N/A s	TITAN RTX @ 1.35 Ghz (Python)
H. O, C. Yang and K. Huh: SeSame: Simple, Easy 3D Object Detection with Point-Wise Semantics. Proceedings of the Asian Conference on Computer Vision (ACCV) 2024.
171	PL++: PV-RCNN++			50.51 %	60.54 %	47.30 %	0.342 s	RTX 4060Ti (Python)

172	DensePointPillars			49.81 %	58.95 %	47.05 %	0.03 s	GPU @ 2.5 Ghz (Python)

173	SeSame-voxel		code	49.74 %	60.69 %	45.64 %	N/A s	TITAN RTX @ 1.35 Ghz (Python)
H. O, C. Yang and K. Huh: SeSame: Simple, Easy 3D Object Detection with Point-Wise Semantics. Proceedings of the Asian Conference on Computer Vision (ACCV) 2024.
174	SCNet			49.61 %	60.95 %	46.91 %	0.04 s	GPU @ 3.0 Ghz (Python)
Z. Wang, H. Fu, L. Wang, L. Xiao and B. Dai: SCNet: Subdivision Coding Network for Object Detection Based on 3D Point Cloud. IEEE Access 2019.
175	MMLab-PointRCNN		code	49.41 %	58.93 %	46.44 %	0.1 s	GPU @ 2.5 Ghz (Python + C/C++)
S. Shi, X. Wang and H. Li: Pointrcnn: 3d object proposal generation and detection from point cloud. Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition 2019.
176	HomoLoss(monoflex)		code	48.97 %	63.77 %	44.60 %	0.04 s	1 core @ 2.5 Ghz (Python)
J. Gu, B. Wu, L. Fan, J. Huang, S. Cao, Z. Xiang and X. Hua: Homography Loss for Monocular 3D Object Detection. Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) 2022.
177	mdab			48.66 %	65.70 %	43.93 %	0.02 s	1 core @ 2.5 Ghz (Python)

178	mdab			48.66 %	65.70 %	43.93 %	22 s	1 core @ 2.5 Ghz (C/C++)

179	ACFD		code	48.63 %	61.62 %	44.15 %	0.2 s	4 cores @ >3.5 Ghz (C/C++)
M. Dimitrievski, P. Veelaert and W. Philips: Semantically aware multilateral filter for depth upsampling in automotive LiDAR point clouds. IEEE Intelligent Vehicles Symposium, IV 2017, Los Angeles, CA, USA, June 11-14, 2017 2017.
180	R-CNN			48.57 %	62.88 %	43.05 %	4 s	GPU @ 3.3 Ghz (C/C++)
J. Hosang, M. Omran, R. Benenson and B. Schiele: Taking a Deeper Look at Pedestrians. arXiv 2015.
181	GraphAlign(ICCV2023)		code	48.47 %	55.17 %	46.68 %	0.03 s	GPU @ 2.0 Ghz (Python)
Z. Song, H. Wei, L. Bai, L. Yang and C. Jia: GraphAlign: Enhancing accurate feature alignment by graph matching for multi-modal 3D object detection. Proceedings of the IEEE/CVF International Conference on Computer Vision 2023.
182	Fade-kd			47.74 %	58.54 %	45.51 %	0.01 s	1 core @ 2.5 Ghz (C/C++)

183	MonoLiG		code	47.69 %	62.87 %	43.27 %	0.03 s	1 core @ 2.5 Ghz (C/C++)
A. Hekimoglu, M. Schmidt and A. Ramiro: Monocular 3D Object Detection with LiDAR Guided Semi Supervised Active Learning. 2023.
184	focal			47.64 %	55.88 %	45.66 %	100 s	1 core @ 2.5 Ghz (Python)

185	MonoFlex			47.58 %	62.64 %	43.15 %	0.03 s	GPU @ 2.5 Ghz (Python)
Y. Zhang, J. Lu and J. Zhou: Objects are Different: Flexible Monocular 3D Object Detection. CVPR 2021.
186	BirdNet+		code	47.50 %	54.78 %	45.53 %	0.11 s	Titan Xp (PyTorch)
A. Barrera, J. Beltrán, C. Guindel, J. Iglesias and F. García: BirdNet+: Two-Stage 3D Object Detection in LiDAR through a Sparsity-Invariant Bird’s Eye View. IEEE Access 2021.
187	geo-pillars			46.92 %	55.56 %	44.74 %	0.05 s	1 core @ 2.5 Ghz (C/C++)
ERROR: Wrong syntax in BIBTEX file.
188	CMKD		code	46.84 %	61.04 %	42.92 %	0.1 s	1 core @ 2.5 Ghz (C/C++)
Y. Hong, H. Dai and Y. Ding: Cross-Modality Knowledge Distillation Network for Monocular 3D Object Detection. ECCV 2022.
189	MonOAPC			46.31 %	60.93 %	42.05 %	0035 s	1 core @ 2.5 Ghz (Python)
H. Yao, J. Chen, Z. Wang, X. Wang, P. Han, X. Chai and Y. Qiu: Occlusion-Aware Plane-Constraints for Monocular 3D Object Detection. IEEE Transactions on Intelligent Transportation Systems 2023.
190	SS3D			45.79 %	61.58 %	41.14 %	48 ms	Tesla V100 (Python)
E. Jörgensen, C. Zach and F. Kahl: Monocular 3D Object Detection and Box Fitting Trained End-to-End Using Intersection-over-Union Loss. CoRR 2019.
191	MonoRCNN++		code	45.76 %	60.29 %	39.39 %	0.07 s	GPU @ 2.5 Ghz (Python)
X. Shi, Z. Chen and T. Kim: Multivariate Probabilistic Monocular 3D Object Detection. WACV 2023.
192	ACF			45.67 %	59.81 %	40.88 %	1 s	1 core @ 3.5 Ghz (Matlab + C/C++)
P. Doll\'ar, R. Appel, S. Belongie and P. Perona: Fast Feature Pyramids for Object Detection. PAMI 2014.
193	UniCuboid			45.64 %	59.74 %	41.29 %	0.1 s	1 core @ 2.5 Ghz (C/C++)

194	Fusion-DPM		code	44.99 %	58.93 %	40.19 %	~ 30 s	1 core @ 3.5 Ghz (Matlab + C/C++)
C. Premebida, J. Carreira, J. Batista and U. Nunes: Pedestrian Detection Combining RGB and Dense LIDAR Data. IROS 2014.
195	ACF-MR			44.79 %	58.29 %	39.94 %	0.6 s	1 core @ 3.5 Ghz (C/C++)
R. Rajaram, E. Ohn-Bar and M. Trivedi: Looking at Pedestrians at Different Scales: A Multi-resolution Approach and Evaluations. T-ITS 2016.
196	SeSame-pillar		code	44.21 %	52.67 %	41.95 %	N/A s	TITAN RTX @ 1.35 Ghz (Python)
H. O, C. Yang and K. Huh: SeSame: Simple, Easy 3D Object Detection with Point-Wise Semantics. Proceedings of the Asian Conference on Computer Vision (ACCV) 2024.
197	LPCG-Monoflex		code	44.13 %	62.44 %	39.46 %	0.03 s	1 core @ 2.5 Ghz (C/C++)
L. Peng, F. Liu, Z. Yu, S. Yan, D. Deng, Z. Yang, H. Liu and D. Cai: Lidar Point Cloud Guided Monocular 3D Object Detection. ECCV 2022.
198	HA-SSVM			43.87 %	58.76 %	38.81 %	21 s	1 core @ >3.5 Ghz (Matlab + C/C++)
J. Xu, S. Ramos, D. Vázquez and A. López: Hierarchical Adaptive Structural SVM for Domain Adaptation. IJCV 2016.
199	AB3DMOT		code	43.86 %	54.55 %	40.99 %	0.0047s	1 core @ 2.5 Ghz (python)
X. Weng and K. Kitani: A Baseline for 3D Multi-Object Tracking. arXiv:1907.03961 2019.
200	MonoEF			43.73 %	58.79 %	39.45 %	0.03 s	1 core @ 2.5 Ghz (Python)
Y. Zhou, Y. He, H. Zhu, C. Wang, H. Li and Q. Jiang: Monocular 3D Object Detection: An Extrinsic Parameter Free Approach. Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) 2021.
201	D4LCN		code	43.50 %	59.55 %	37.12 %	0.2 s	GPU @ 2.5 Ghz (Python + C/C++)
M. Ding, Y. Huo, H. Yi, Z. Wang, J. Shi, Z. Lu and P. Luo: Learning Depth-Guided Convolutions for Monocular 3D Object Detection. CVPR 2020.
202	DMF			43.43 %	52.99 %	41.29 %	0.2 s	1 core @ 2.5 Ghz (Python + C/C++)
X. J. Chen and W. Xu: Disparity-Based Multiscale Fusion Network for Transportation Detection. IEEE Transactions on Intelligent Transportation Systems 2022.
203	MonoDDE			43.36 %	57.80 %	39.00 %	0.04 s	1 core @ 2.5 Ghz (Python)
Z. Li, Z. Qu, Y. Zhou, J. Liu, H. Wang and L. Jiang: Diversity Matters: Fully Exploiting Depth Clues for Reliable Monocular 3D Object Detection. CVPR 2022.
204	T-SSD			43.31 %	53.86 %	40.96 %	0.04	1 core @ 2.0 Ghz (C/C++)

205	DPM-VOC+VP			43.26 %	59.21 %	38.12 %	8 s	1 core @ 2.5 Ghz (C/C++)
B. Pepik, M. Stark, P. Gehler and B. Schiele: Multi-view and 3D Deformable Part Models. IEEE Transactions on Pattern Analysis and Machine Intelligence (TPAMI) 2015.
206	ACF-SC			42.97 %	53.30 %	38.12 %	<0.3 s	1 core @ >3.5 Ghz (Matlab + C/C++)
C. Cadena, A. Dick and I. Reid: A Fast, Modular Scene Understanding System using Context-Aware Object Detection. Robotics and Automation (ICRA), 2015 IEEE International Conference on 2015.
207	SeSame-voxel w/score		code	42.88 %	50.84 %	40.76 %	N/A s	GPU @ 1.5 Ghz (Python)
H. O, C. Yang and K. Huh: SeSame: Simple, Easy 3D Object Detection with Point-Wise Semantics. Proceedings of the Asian Conference on Computer Vision (ACCV) 2024.
208	MonoDTR			42.86 %	59.44 %	38.57 %	0.04 s	1 core @ 2.5 Ghz (C/C++)
K. Huang, T. Wu, H. Su and W. Hsu: MonoDTR: Monocular 3D Object Detection with Depth-Aware Transformer. CVPR 2022.
209	SquaresICF		code	42.61 %	57.08 %	37.85 %	1 s	GPU @ >3.5 Ghz (C/C++)
R. Benenson, M. Mathias, T. Tuytelaars and L. Gool: Seeking the strongest rigid detector. CVPR 2013.
210	CG-Stereo			42.54 %	54.64 %	38.45 %	0.57 s	GeForce RTX 2080 Ti
C. Li, J. Ku and S. Waslander: Confidence Guided Stereo 3D Object Detection with Split Depth Estimation. IROS 2020.
211	BirdNet+ (legacy)		code	41.97 %	51.38 %	40.15 %	0.1 s	Titan Xp (PyTorch)
A. Barrera, C. Guindel, J. Beltrán and F. García: BirdNet+: End-to-End 3D Object Detection in LiDAR Bird’s Eye View. 2020 IEEE 23rd International Conference on Intelligent Transportation Systems (ITSC) 2020.
212	DDMP-3D			41.54 %	56.73 %	35.52 %	0.18 s	1 core @ 2.5 Ghz (Python)
L. Wang, L. Du, X. Ye, Y. Fu, G. Guo, X. Xue, J. Feng and L. Zhang: Depth-conditioned Dynamic Message Propagation for Monocular 3D Object Detection. CVPR 2020.
213	CSW3D			41.50 %	53.76 %	37.25 %	0.03 s	4 cores @ 2.5 Ghz (C/C++)
J. Hu, T. Wu, H. Fu, Z. Wang and K. Ding: Cascaded Sliding Window Based Real-Time 3D Region Proposal for Pedestrian Detection. ROBIO 2019.
214	M3D-RPN		code	41.46 %	56.64 %	37.31 %	0.16 s	GPU @ 1.5 Ghz (Python)
G. Brazil and X. Liu: M3D-RPN: Monocular 3D Region Proposal Network for Object Detection . ICCV 2019 .
215	YOLOStereo3D		code	41.46 %	56.20 %	37.07 %	0.1 s	GPU 1080Ti
Y. Liu, L. Wang and M. Liu: YOLOStereo3D: A Step Back to 2D for Efficient Stereo 3D Detection. 2021 International Conference on Robotics and Automation (ICRA) 2021.
216	MonoFRD			41.20 %	54.06 %	37.53 %	0.01 s	1 core @ 2.5 Ghz (C/C++)
Z. Gong, Y. Zhao, F. Zhang, G. Gui, B. Chen, L. Yu, H. Wang, C. Yang and W. Gui: Color intuitive feature guided depth-height fusion and volume rendering for monocular 3D object detection. IEEE Transactions on Intelligent Vehicles(Major Revison) 2024.
217	CIE			41.04 %	53.27 %	37.73 %	0.1 s	1 core @ 2.5 Ghz (C/C++)
Anonymities: Consistency of Implicit and Explicit Features Matters for Monocular 3D Object Detection. arXiv preprint arXiv:2207.07933 2022.
218	SubCat			40.50 %	53.75 %	35.66 %	1.2 s	6 cores @ 2.5 Ghz (Matlab + C/C++)
E. Ohn-Bar and M. Trivedi: Fast and Robust Object Detection Using Visual Subcategories. Computer Vision and Pattern Recognition Workshops Mobile Vision 2014.
219	PS-fld		code	40.47 %	55.47 %	36.65 %	0.25 s	1 core @ 2.5 Ghz (C/C++)
Y. Chen, H. Dai and Y. Ding: Pseudo-Stereo for Monocular 3D Object Detection in Autonomous Driving. Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) 2022.
220	SeSame-pillar w/scor		code	40.24 %	48.38 %	38.25 %	N/A s	1 core @ 2.5 Ghz (C/C++)
H. O, C. Yang and K. Huh: SeSame: Simple, Easy 3D Object Detection with Point-Wise Semantics. Proceedings of the Asian Conference on Computer Vision (ACCV) 2024.
221	DSGN		code	39.93 %	49.28 %	38.13 %	0.67 s	NVIDIA Tesla V100
Y. Chen, S. Liu, X. Shen and J. Jia: DSGN: Deep Stereo Geometry Network for 3D Object Detection. CVPR 2020.
222	RT3D-GMP			39.83 %	55.56 %	35.18 %	0.06 s	GPU @ 2.5 Ghz (Python + C/C++)
H. Königshof and C. Stiller: Learning-Based Shape Estimation with Grid Map Patches for Realtime 3D Object Detection for Automated Driving. 2020 IEEE 23rd International Conference on Intelligent Transportation Systems (ITSC) 2020.
223	SparsePool		code	39.59 %	50.81 %	35.91 %	0.13 s	8 cores @ 2.5 Ghz (Python)
Z. Wang, W. Zhan and M. Tomizuka: Fusing bird view lidar point cloud and front view camera image for deep object detection. arXiv preprint arXiv:1711.06703 2017.
224	SparsePool		code	39.43 %	50.94 %	35.78 %	0.13 s	8 cores @ 2.5 Ghz (Python)
Z. Wang, W. Zhan and M. Tomizuka: Fusing bird view lidar point cloud and front view camera image for deep object detection. arXiv preprint arXiv:1711.06703 2017.
225	AVOD		code	39.43 %	50.90 %	35.75 %	0.08 s	Titan X (pascal)
J. Ku, M. Mozifian, J. Lee, A. Harakeh and S. Waslander: Joint 3D Proposal Generation and Object Detection from View Aggregation. IROS 2018.
226	ACF			39.12 %	48.42 %	35.03 %	0.2 s	1 core @ >3.5 Ghz (Matlab + C/C++)
P. Doll\'ar, R. Appel, S. Belongie and P. Perona: Fast Feature Pyramids for Object Detection. PAMI 2014. P. Doll\'ar: Piotr's Image and Video Matlab Toolbox (PMT). .
227	LSVM-MDPM-sv			37.26 %	50.74 %	33.13 %	10 s	4 cores @ 3.0 Ghz (C/C++)
P. Felzenszwalb, R. Girshick, D. McAllester and D. Ramanan: Object Detection with Discriminatively Trained Part-Based Models. PAMI 2010. A. Geiger, C. Wojek and R. Urtasun: Joint 3D Estimation of Objects and Scene Layout. NIPS 2011.
228	multi-task CNN			37.00 %	49.38 %	33.46 %	25.1 ms	GPU @ 2.0 Ghz (Python)
M. Oeljeklaus, F. Hoffmann and T. Bertram: A Fast Multi-Task CNN for Spatial Understanding of Traffic Scenes. IEEE Intelligent Transportation Systems Conference 2018.
229	Complexer-YOLO			36.45 %	42.16 %	32.91 %	0.06 s	GPU @ 3.5 Ghz (C/C++)
M. Simon, K. Amende, A. Kraus, J. Honer, T. Samann, H. Kaulbersch, S. Milz and H. Michael Gross: Complexer-YOLO: Real-Time 3D Object Detection and Tracking on Semantic Point Clouds. The IEEE Conference on Computer Vision and Pattern Recognition (CVPR) Workshops 2019.
230	LSVM-MDPM-us		code	35.92 %	48.73 %	31.70 %	10 s	4 cores @ 3.0 Ghz (C/C++)
P. Felzenszwalb, R. Girshick, D. McAllester and D. Ramanan: Object Detection with Discriminatively Trained Part-Based Models. PAMI 2010.
231	CMAN			34.96 %	49.73 %	30.92 %	0.15 s	1 core @ 2.5 Ghz (Python)
C. Yuanzhouhan Cao: CMAN: Leaning Global Structure Correlation for Monocular 3D Object Detection. IEEE Trans. Intell. Transport. Syst. 2022.
232	Aug3D-RPN			34.95 %	47.22 %	30.64 %	0.08 s	1 core @ 2.5 Ghz (C/C++)
C. He, J. Huang, X. Hua and L. Zhang: Aug3D-RPN: Improving Monocular 3D Object Detection by Synthetic Images with Virtual Depth. 2021.
233	FMF-occlusion-net			34.74 %	49.26 %	30.37 %	0.16 s	1 core @ 2.5 Ghz (Python + C/C++)
H. Liu, H. Liu, Y. Wang, F. Sun and W. Huang: Fine-grained Multi-level Fusion for Anti- occlusion Monocular 3D Object Detection. IEEE Transactions on Image Processing 2022.
234	Fade			34.70 %	43.64 %	32.98 %	0.01 s	1 core @ 2.5 Ghz (C/C++)

235	MonoNeRD		code	34.43 %	46.50 %	31.06 %	na s	1 core @ 2.5 Ghz (Python)
J. Xu, L. Peng, H. Cheng, H. Li, W. Qian, K. Li, W. Wang and D. Cai: MonoNeRD: NeRF-like Representations for Monocular 3D Object Detection. ICCV 2023.
236	PS-SVDM			34.15 %	46.43 %	30.90 %	1 s	1 core @ 2.5 Ghz (Python)
Y. Shi: SVDM: Single-View Diffusion Model for Pseudo-Stereo 3D Object Detection. arXiv preprint arXiv:2307.02270 2023.
237	LLW			34.15 %	46.50 %	30.83 %	0.1 s	1 core @ 2.5 Ghz (C/C++)

238	PointRGBNet			33.92 %	44.35 %	30.43 %	0.08 s	4 cores @ 2.5 Ghz (Python + C/C++)
P. Xie Desheng: Real-time Detection of 3D Objects Based on Multi-Sensor Information Fusion. Automotive Engineering 2022.
239	PGD-FCOS3D		code	33.67 %	48.30 %	29.76 %	0.03 s	1 core @ 2.5 Ghz (C/C++)
T. Wang, X. Zhu, J. Pang and D. Lin: Probabilistic and Geometric Depth: Detecting Objects in Perspective. Conference on Robot Learning (CoRL) 2021.
240	Vote3D			33.04 %	42.66 %	30.59 %	0.5 s	4 cores @ 2.8 Ghz (C/C++)
D. Wang and I. Posner: Voting for Voting in Online Point Cloud Object Detection. Proceedings of Robotics: Science and Systems 2015.
241	ESGN			32.60 %	44.09 %	29.10 %	0.06 s	GPU @ 2.5 Ghz (Python + C/C++)
A. Gao, Y. Pang, J. Nie, Z. Shao, J. Cao, Y. Guo and X. Li: ESGN: Efficient Stereo Geometry Network for Fast 3D Object Detection. IEEE Transactions on Circuits and Systems for Video Technology 2022.
242	SGM3D		code	32.48 %	45.03 %	28.58 %	0.03 s	1 core @ 2.5 Ghz (C/C++)
Z. Zhou, L. Du, X. Ye, Z. Zou, X. Tan, L. Zhang, X. Xue and J. Feng: SGM3D: Stereo Guided Monocular 3D Object Detection. RA-L 2022.
243	CPD++(unsupervised)		code	32.46 %	39.25 %	30.79 %	0.1 s	GPU @ >3.5 Ghz (Python)

244	CaDDN		code	32.42 %	46.35 %	29.98 %	0.63 s	GPU @ 2.5 Ghz (Python)
C. Reading, A. Harakeh, J. Chae and S. Waslander: Categorical Depth Distribution Network for Monocular 3D Object Detection. CVPR 2021.
245	PS-SVDM			32.24 %	44.02 %	29.08 %	1 s	1 core @ 2.5 Ghz (Python)
Y. Shi: SVDM: Single-View Diffusion Model for Pseudo-Stereo 3D Object Detection. arXiv preprint arXiv:2307.02270 2023.
246	DFR-Net			31.84 %	45.20 %	27.94 %	0.18 s	1080 Ti (Pytorch)
Z. Zou, X. Ye, L. Du, X. Cheng, X. Tan, L. Zhang, J. Feng, X. Xue and E. Ding: The devil is in the task: Exploiting reciprocal appearance-localization features for monocular 3d object detection . ICCV 2021.
247	OC Stereo		code	30.79 %	43.50 %	28.40 %	0.35 s	1 core @ 2.5 Ghz (Python + C/C++)
A. Pon, J. Ku, C. Li and S. Waslander: Object-Centric Stereo Matching for 3D Object Detection. ICRA 2020.
248	mBoW			30.26 %	41.52 %	26.34 %	10 s	1 core @ 2.5 Ghz (C/C++)
J. Behley, V. Steinhage and A. Cremers: Laser-based Segment Classification Using a Mixture of Bag-of-Words. Proc. of the IEEE/RSJ International Conference on Intelligent Robots and Systems (IROS) 2013.
249	BirdNet			30.07 %	36.82 %	28.40 %	0.11 s	Titan Xp (Caffe)
J. Beltrán, C. Guindel, F. Moreno, D. Cruzado, F. García and A. Escalera: BirdNet: A 3D Object Detection Framework from LiDAR Information. 2018 21st International Conference on Intelligent Transportation Systems (ITSC) 2018.
250	SeSame-point w/score		code	30.04 %	40.65 %	27.65 %	N/A s	1 core @ 1.5 Ghz (Python)

	Method	Setting	Code	Moderate	Easy	Hard	Runtime	Environment
1	LCANet			86.37 %	92.08 %	79.01 %	1 s	1 core @ 2.5 Ghz (C/C++)

2	UPIDet		code	84.44 %	90.16 %	77.71 %	0.11 s	1 core @ 2.5 Ghz (C/C++)
Y. Zhang, Q. Zhang, J. Hou, Y. Yuan and G. Xing: Unleash the Potential of Image Branch for Cross-modal 3D Object Detection. Thirty-seventh Conference on Neural Information Processing Systems 2023.
3	TED		code	84.36 %	92.60 %	78.43 %	0.1 s	1 core @ 2.5 Ghz (C/C++)
H. Wu, C. Wen, W. Li, R. Yang and C. Wang: Transformation-Equivariant 3D Object Detection for Autonomous Driving. AAAI 2023.
4	CasA++		code	84.26 %	92.38 %	78.42 %	0.1 s	1 core @ 2.5 Ghz (C/C++)
H. Wu, J. Deng, C. Wen, X. Li and C. Wang: CasA: A Cascade Attention Network for 3D Object Detection from LiDAR point clouds. IEEE Transactions on Geoscience and Remote Sensing 2022.
5	LoGoNet		code	84.00 %	90.14 %	77.97 %	0.1 s	1 core @ 2.5 Ghz (C/C++)
X. Li, T. Ma, Y. Hou, B. Shi, Y. Yang, Y. Liu, X. Wu, Q. Chen, Y. Li, Y. Qiao and others: LoGoNet: Towards Accurate 3D Object Detection with Local-to-Global Cross-Modal Fusion. CVPR 2023.
6	CasA		code	83.21 %	92.86 %	77.12 %	0.1 s	1 core @ 2.5 Ghz (C/C++)
H. Wu, J. Deng, C. Wen, X. Li and C. Wang: CasA: A Cascade Attention Network for 3D Object Detection from LiDAR point clouds. IEEE Transactions on Geoscience and Remote Sensing 2022.
7	MLF-DET			81.95 %	87.34 %	74.79 %	0.09 s	1 core @ 2.5 Ghz (C/C++)
Z. Lin, Y. Shen, S. Zhou, S. Chen and N. Zheng: MLF-DET: Multi-Level Fusion for Cross- Modal 3D Object Detection. International Conference on Artificial Neural Networks 2023.
8	HMFI		code	81.76 %	89.35 %	74.93 %	0.1 s	1 core @ 2.5 Ghz (C/C++)
X. Li, B. Shi, Y. Hou, X. Wu, T. Ma, Y. Li and L. He: Homogeneous Multi-modal Feature Fusion and Interaction for 3D Object Detection. ECCV 2022.
9	MPC3DNet			81.73 %	88.19 %	75.09 %	0.05 s	GPU @ 1.5 Ghz (Python)

10	RangeIoUDet			81.67 %	90.43 %	74.90 %	0.02 s	GPU @ 2.5 Ghz (Python + C/C++)
Z. Liang, Z. Zhang, M. Zhang, X. Zhao and S. Pu: RangeIoUDet: Range Image Based Real-Time 3D Object Detector Optimized by Intersection Over Union. CVPR 2021.
11	USVLab BSAODet		code	81.36 %	86.82 %	74.40 %	0.04 s	1 core @ 2.5 Ghz (C/C++)
W. Xiao, Y. Peng, C. Liu, J. Gao, Y. Wu and X. Li: Balanced Sample Assignment and Objective for Single-Model Multi-Class 3D Object Detection. IEEE Transactions on Circuits and Systems for Video Technology 2023.
12	RagNet3D		code	81.20 %	90.12 %	74.57 %	0.05 s	1 core @ 2.5 Ghz (C/C++)
J. Chen, Y. Han, Z. Yan, J. Qian, J. Li and J. Yang: Ragnet3d: Learning Distinguishable Representation for Pooled Grids in 3d Object Detection. Available at SSRN 4979473 .
13	AFFN			81.03 %	89.88 %	75.05 %	0.5 s	1 core @ 2.5 Ghz (C/C++)

14	GEFPN			81.03 %	89.88 %	75.05 %	0.5 s	1 core @ 2.5 Ghz (C/C++)

15	CAT-Det			80.70 %	87.94 %	73.86 %	0.3 s	GPU @ 2.5 Ghz (Python + C/C++)
Y. Zhang, J. Chen and D. Huang: CAT-Det: Contrastively Augmented Transformer for Multi-modal 3D Object Detection. CVPR 2022.
16	SPG_mini		code	80.58 %	87.77 %	74.86 %	0.09 s	GPU @ 2.5 Ghz (Python)
Q. Xu, Y. Zhou, W. Wang, C. Qi and D. Anguelov: SPG: Unsupervised Domain Adaptation for 3D Object Detection via Semantic Point Generation. Proceedings of the IEEE conference on computer vision and pattern recognition (ICCV) 2021.
17	DSA-PV-RCNN		code	80.57 %	88.65 %	74.81 %	0.08 s	1 core @ 2.5 Ghz (Python + C/C++)
P. Bhattacharyya, C. Huang and K. Czarnecki: SA-Det3D: Self-Attention Based Context-Aware 3D Object Detection. 2021.
18	BtcDet		code	80.46 %	88.41 %	74.59 %	0.09 s	GPU @ 2.5 Ghz (Python + C/C++)
Q. Xu, Y. Zhong and U. Neumann: Behind the Curtain: Learning Occluded Shapes for 3D Object Detection. Proceedings of the AAAI Conference on Artificial Intelligence 2022.
19	LPFusion_three_class			80.43 %	87.56 %	71.44 %	0.1 s	1 core @ 2.5 Ghz (C/C++)

20	MMLab PV-RCNN		code	80.42 %	86.62 %	73.64 %	0.08 s	1 core @ 2.5 Ghz (Python + C/C++)
S. Shi, C. Guo, L. Jiang, Z. Wang, J. Shi, X. Wang and H. Li: PV-RCNN: Point-Voxel Feature Set Abstraction for 3D Object Detection. CVPR 2020.
21	EQ-PVRCNN		code	80.37 %	89.07 %	74.20 %	0.2 s	GPU @ 2.5 Ghz (Python + C/C++)
Z. Yang, L. Jiang, Y. Sun, B. Schiele and J. Jia: A Unified Query-based Paradigm for Point Cloud Understanding. Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition 2022.
22	test			80.18 %	91.30 %	73.97 %	0.04 s	GPU @ 1.5 Ghz (Python + C/C++)

23	HINTED		code	80.04 %	86.76 %	73.45 %	0.04 s	1 core @ 2.5 Ghz (C/C++)
Q. Xia, W. Ye, H. Wu, S. Zhao, L. Xing, X. Huang, J. Deng, X. Li, C. Wen and C. Wang: HINTED: Hard Instance Enhanced Detector with Mixed-Density Feature Fusion for Sparsely- Supervised 3D Object Detection. Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition 2024.
24	PDV		code	79.84 %	88.76 %	73.04 %	0.1 s	1 core @ 2.5 Ghz (C/C++)
J. Hu, T. Kuai and S. Waslander: Point Density-Aware Voxels for LiDAR 3D Object Detection. CVPR 2022.
25	SCNet3D			79.78 %	89.85 %	72.85 %	0.08 s	1 core @ 2.5 Ghz (Python)

26	vsis-PHNet			79.54 %	90.67 %	74.37 %	0.03 s	1 core @ 2.5 Ghz (C/C++)

27	M3DeTR		code	79.29 %	87.38 %	72.46 %	n/a s	GPU @ 1.0 Ghz (Python)
T. Guan, J. Wang, S. Lan, R. Chandra, Z. Wu, L. Davis and D. Manocha: M3DeTR: Multi-representation, Multi- scale, Mutual-relation 3D Object Detection with Transformers. 2021.
28	PASS-PV-RCNN-Plus			79.22 %	86.26 %	72.68 %	1 s	1 core @ 2.5 Ghz (Python)
Anonymous: Leveraging Anchor-based LiDAR 3D Object Detection via Point Assisted Sample Selection. will submit to computer vision conference/journal 2024.
29	BVIFusion+			79.18 %	88.31 %	72.49 %	0.09 s	1 core @ 2.5 Ghz (Python)

30	HotSpotNet			78.81 %	86.06 %	71.74 %	0.04 s	1 core @ 2.5 Ghz (Python + C/C++)
Q. Chen, L. Sun, Z. Wang, K. Jia and A. Yuille: object as hotspots. Proceedings of the European Conference on Computer Vision (ECCV) 2020.
31	IA-SSD (single)		code	78.71 %	88.99 %	72.03 %	0.013 s	1 core @ 2.5 Ghz (C/C++)
Y. Zhang, Q. Hu, G. Xu, Y. Ma, J. Wan and Y. Guo: Not All Points Are Equal: Learning Highly Efficient Point-based Detectors for 3D LiDAR Point Clouds. CVPR 2022.
32	PG-RCNN		code	78.69 %	88.17 %	72.16 %	0.06 s	GPU @ 1.5 Ghz (Python)
I. Koo, I. Lee, S. Kim, H. Kim, W. Jeon and C. Kim: PG-RCNN: Semantic Surface Point Generation for 3D Object Detection. 2023.
33	PHNetc			78.57 %	89.35 %	73.56 %	0.03 s	1 core @ 2.5 Ghz (C/C++)

34	AFFN-G			78.29 %	85.19 %	73.15 %	0.5 s	1 core @ 2.5 Ghz (C/C++)

35	MMLab-PartA^2		code	78.29 %	88.90 %	71.19 %	0.08 s	GPU @ 2.5 Ghz (Python + C/C++)
S. Shi, Z. Wang, J. Shi, X. Wang and H. Li: From Points to Parts: 3D Object Detection from Point Cloud with Part-aware and Part-aggregation Network. IEEE Transactions on Pattern Analysis and Machine Intelligence 2020.
36	F-ConvNet		code	78.05 %	86.75 %	68.12 %	0.47 s	GPU @ 2.5 Ghz (Python + C/C++)
Z. Wang and K. Jia: Frustum ConvNet: Sliding Frustums to Aggregate Local Point-Wise Features for Amodal 3D Object Detection. IROS 2019.
37	PointPainting			78.04 %	87.70 %	69.27 %	0.4 s	GPU @ 2.5 Ghz (Python + C/C++)
S. Vora, A. Lang, B. Helou and O. Beijbom: PointPainting: Sequential Fusion for 3D Object Detection. CVPR 2020.
38	dsvd+vx			77.89 %	91.96 %	70.70 %	0.1 s	1 core @ 2.5 Ghz (C/C++)

39	FIRM-Net_SCF+			77.82 %	89.67 %	70.82 %	0.07 s	1 core @ 2.5 Ghz (Python)

40	DFAF3D			77.74 %	87.20 %	70.77 %	0.05 s	1 core @ 2.5 Ghz (Python)
Q. Tang, X. Bai, J. Guo, B. Pan and W. Jiang: DFAF3D: A dual-feature-aware anchor-free single-stage 3D detector for point clouds. Image and Vision Computing 2023.
41	MFB3D			77.67 %	86.23 %	71.68 %	0.14 s	1 core @ 2.5 Ghz (Python)

42	OGMMDet		code	77.59 %	86.71 %	72.35 %	0.01 s	1 core @ 2.5 Ghz (C/C++)

43	ANM		code	77.59 %	88.61 %	72.35 %	ANM	ANM

44	FIRM-Net-SCF			77.58 %	89.59 %	70.62 %	0.07 s	1 core @ 2.5 Ghz (Python)

45	CGML			77.49 %	85.87 %	72.13 %	0.33 s	1 core @ 2.5 Ghz (C/C++)

46	3ONet			77.36 %	89.11 %	70.31 %	0.1 s	1 core @ 2.5 Ghz (C/C++)
H. Hoang and M. Yoo: 3ONet: 3-D Detector for Occluded Object Under Obstructed Conditions. IEEE Sensors Journal 2023.
47	GraphAlign(ICCV2023)		code	77.15 %	84.72 %	72.34 %	0.03 s	GPU @ 2.0 Ghz (Python)
Z. Song, H. Wei, L. Bai, L. Yang and C. Jia: GraphAlign: Enhancing accurate feature alignment by graph matching for multi-modal 3D object detection. Proceedings of the IEEE/CVF International Conference on Computer Vision 2023.
48	AFFN-Ga			77.04 %	84.81 %	71.13 %	0.5 s	1 core @ 2.5 Ghz (C/C++)

49	P2V-RCNN			76.93 %	88.40 %	70.35 %	0.1 s	2 cores @ 2.5 Ghz (Python)
J. Li, S. Luo, Z. Zhu, H. Dai, A. Krylov, Y. Ding and L. Shao: P2V-RCNN: Point to Voxel Feature Learning for 3D Object Detection from Point Clouds. IEEE Access 2021.
50	EOTL		code	76.88 %	85.62 %	66.04 %	TBD s	1 core @ 2.5 Ghz (Python + C/C++)
R. Yang, Z. Yan, T. Yang, Y. Wang and Y. Ruichek: Efficient Online Transfer Learning for Road Participants Detection in Autonomous Driving. IEEE Sensors Journal 2023.
51	RRC		code	76.81 %	86.81 %	66.59 %	3.6 s	GPU @ 2.5 Ghz (C/C++)
J. Ren, X. Chen, J. Liu, W. Sun, J. Pang, Q. Yan, Y. Tai and L. Xu: Accurate Single Stage Detector Using Recurrent Rolling Convolution. CVPR 2017.
52	SA V1			76.44 %	87.06 %	69.68 %	0.5 s	GPU @ 2.5 Ghz (Python)

53	CG-SSD			76.25 %	85.36 %	70.09 %	0.01 s	1 core @ 2.5 Ghz (C/C++)

54	ACFNet			76.15 %	86.92 %	71.33 %	0.11 s	1 core @ 2.5 Ghz (C/C++)
Y. Tian, X. Zhang, X. Wang, J. Xu, J. Wang, R. Ai, W. Gu and W. Ding: ACF-Net: Asymmetric Cascade Fusion for 3D Detection With LiDAR Point Clouds and Images. IEEE Transactions on Intelligent Vehicles 2023.
55	focal			76.02 %	84.14 %	70.24 %	100 s	1 core @ 2.5 Ghz (Python)

56	GeVo			76.02 %	84.14 %	70.24 %	0.05 s	1 core @ 2.5 Ghz (C/C++)

57	CAIA_PRO		code	75.98 %	86.78 %	69.51 %	0.01 s	1 core @ 2.5 Ghz (C/C++)

58	ACDet		code	75.41 %	88.54 %	69.45 %	0.05 s	1 core @ 2.5 Ghz (C/C++)
J. Xu, G. Wang, X. Zhang and G. Wan: ACDet: Attentive Cross-view Fusion for LiDAR-based 3D Object Detection. 3DV 2022.
59	MLFusion-VS			75.35 %	83.72 %	69.56 %	0.06 s	1 core @ 2.5 Ghz (C/C++)

60	MS-CNN		code	75.30 %	84.88 %	65.27 %	0.4 s	GPU @ 2.5 Ghz (C/C++)
Z. Cai, Q. Fan, R. Feris and N. Vasconcelos: A Unified Multi-scale Deep Convolutional Neural Network for Fast Object Detection. ECCV 2016.
61	TuSimple		code	75.22 %	83.68 %	65.22 %	1.6 s	GPU @ 2.5 Ghz (Python + C/C++)
F. Yang, W. Choi and Y. Lin: Exploit all the layers: Fast and accurate cnn object detector with scale dependent pooling and cascaded rejection classifiers. Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition 2016. K. He, X. Zhang, S. Ren and J. Sun: Deep residual learning for image recognition. Proceedings of the IEEE conference on computer vision and pattern recognition 2016.
62	SVGA-Net			75.14 %	85.13 %	68.14 %	0.03s	1 core @ 2.5 Ghz (Python + C/C++)
Q. He, Z. Wang, H. Zeng, Y. Zeng and Y. Liu: SVGA-Net: Sparse Voxel-Graph Attention Network for 3D Object Detection from Point Clouds. AAAI 2022.
63	Point-GNN		code	75.08 %	85.75 %	68.69 %	0.6 s	GPU @ 2.5 Ghz (Python)
W. Shi and R. Rajkumar: Point-GNN: Graph Neural Network for 3D Object Detection in a Point Cloud. CVPR 2020.
64	Fast-CLOCs			75.07 %	89.73 %	67.93 %	0.1 s	GPU @ 2.5 Ghz (Python)
S. Pang, D. Morris and H. Radha: Fast-CLOCs: Fast Camera-LiDAR Object Candidates Fusion for 3D Object Detection. Proceedings of the IEEE/CVF Winter Conference on Applications of Computer Vision (WACV) 2022.
65	GSG-FPS		code	75.03 %	87.60 %	68.61 %	0.01 s	1 core @ 2.5 Ghz (C/C++)

66	SVFMamba		code	75.00 %	83.71 %	68.71 %	N/A s	1 core @ 2.5 Ghz (C/C++)

67	R50_SACINet			74.96 %	86.69 %	68.92 %	0.06 s	1 core @ 2.5 Ghz (C/C++)

68	Voxel RCNN-Focal*		code	74.84 %	81.72 %	69.66 %	0.2 s	1 core @ 2.5 Ghz (Python)

69	Res3DNet			74.82 %	88.67 %	68.19 %	0.05 s	GPU @ 3.5 Ghz (Python)

70	Deep3DBox			74.78 %	84.36 %	64.05 %	1.5 s	GPU @ 2.5 Ghz (C/C++)
A. Mousavian, D. Anguelov, J. Flynn and J. Kosecka: 3D Bounding Box Estimation Using Deep Learning and Geometry. CVPR 2017.
71	LumiNet		code	74.76 %	88.45 %	67.89 %	0.1 s	1 core @ 2.5 Ghz (Python)

72	VPFNet		code	74.52 %	82.60 %	66.04 %	0.2 s	1 core @ 2.5 Ghz (C/C++)
C. Wang, H. Chen and L. Fu: VPFNet: Voxel-Pixel Fusion Network for Multi-class 3D Object Detection. 2021. C. Wang, H. Chen, Y. Chen, P. Hsiao and L. Fu: VoPiFNet: Voxel-Pixel Fusion Network for Multi-Class 3D Object Detection. IEEE Transactions on Intelligent Transportation Systems 2024.
73	L_SACINet			74.49 %	84.34 %	68.25 %	0.04 s	1 core @ 2.5 Ghz (C/C++)

74	3DSSD		code	74.12 %	87.09 %	67.67 %	0.04 s	GPU @ 2.5 Ghz (Python + C/C++)
Z. Yang, Y. Sun, S. Liu and J. Jia: 3DSSD: Point-based 3D Single Stage Object Detector. CVPR 2020.
75	SFA_IGCL_Focalsconv*		code	74.06 %	80.50 %	68.54 %	0.2 s	1 core @ 2.5 Ghz (C/C++)

76	FocalsConv*			74.03 %	83.55 %	69.37 %	0.2 s	1 core @ 2.5 Ghz (C/C++)

77	MSMA V1			74.02 %	85.61 %	67.70 %	0.5 s	GPU @ 2.5 Ghz (Python)

78	SDP+RPN			73.85 %	82.59 %	64.87 %	0.4 s	GPU @ 2.5 Ghz (Python + C/C++)
F. Yang, W. Choi and Y. Lin: Exploit All the Layers: Fast and Accurate CNN Object Detector with Scale Dependent Pooling and Cascaded Rejection Classifiers. Proceedings of the IEEE International Conference on Computer Vision and Pattern Recognition 2016. S. Ren, K. He, R. Girshick and J. Sun: Faster R-CNN: Towards real-time object detection with region proposal networks. Advances in Neural Information Processing Systems 2015.
79	SRDL			73.68 %	85.44 %	66.94 %	0.05 s	1 core @ 2.5 Ghz (Python + C/C++)
ERROR: Wrong syntax in BIBTEX file.
80	DVFENet			73.66 %	85.45 %	67.10 %	0.05 s	1 core @ 2.5 Ghz (Python + C/C++)
Y. He, G. Xia, Y. Luo, L. Su, Z. Zhang, W. Li and P. Wang: DVFENet: Dual-branch Voxel Feature Extraction Network for 3D Object Detection. Neurocomputing 2021.
81	Faraway-Frustum		code	73.63 %	85.43 %	66.64 %	0.1 s	GPU @ 2.5 Ghz (Python)
H. Zhang, D. Yang, E. Yurtsever, K. Redmill and U. Ozguner: Faraway-frustum: Dealing with lidar sparsity for 3D object detection using fusion. 2021 IEEE International Intelligent Transportation Systems Conference (ITSC) 2021.
82	sensekitti		code	73.48 %	82.90 %	64.03 %	4.5 s	GPU @ 2.5 Ghz (Python + C/C++)
B. Yang, J. Yan, Z. Lei and S. Li: Craft Objects from Images. CVPR 2016.
83	L-AUG			73.43 %	83.88 %	68.12 %	0.1 s	1 core @ 2.5 Ghz (C/C++)
T. Cortinhal, I. Gouigah and E. Aksoy: Semantics-aware LiDAR-Only Pseudo Point Cloud Generation for 3D Object Detection. 2023.
84	MMLab-PointRCNN		code	73.42 %	86.21 %	66.45 %	0.1 s	GPU @ 2.5 Ghz (Python + C/C++)
S. Shi, X. Wang and H. Li: Pointrcnn: 3d object proposal generation and detection from point cloud. Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition 2019.
85	SecAtten			73.39 %	83.23 %	67.57 %	0.1 s	1 core @ 2.5 Ghz (C/C++)

86	SIF			73.19 %	85.18 %	65.41 %	0.1 s	1 core @ 2.5 Ghz (C/C++)
P. An: SIF. Submitted to CVIU 2021.
87	F-PointNet		code	73.16 %	86.86 %	65.21 %	0.17 s	GPU @ 3.0 Ghz (Python)
C. Qi, W. Liu, C. Wu, H. Su and L. Guibas: Frustum PointNets for 3D Object Detection from RGB-D Data. arXiv preprint arXiv:1711.08488 2017.
88	FromVoxelToPoint		code	73.16 %	87.07 %	65.98 %	0.1 s	1 core @ 2.5 Ghz (C/C++)
J. Li, H. Dai, L. Shao and Y. Ding: From Voxel to Point: IoU-guided 3D Object Detection for Point Cloud with Voxel-to- Point Decoder. MM '21: The 29th ACM International Conference on Multimedia (ACM MM) 2021.
89	XView			73.16 %	88.02 %	65.37 %	0.1 s	1 core @ 2.5 Ghz (C/C++)
L. Xie, G. Xu, D. Cai and X. He: X-view: Non-egocentric Multi-View 3D Object Detector. 2021.
90	S-AT GCN			72.81 %	82.79 %	66.72 %	0.02 s	GPU @ 2.0 Ghz (Python)
L. Wang, C. Wang, X. Zhang, T. Lan and J. Li: S-AT GCN: Spatial-Attention Graph Convolution Network based Feature Enhancement for 3D Object Detection. CoRR 2021.
91	H^23D R-CNN		code	72.73 %	85.50 %	65.81 %	0.03 s	1 core @ 2.5 Ghz (C/C++)
J. Deng, W. Zhou, Y. Zhang and H. Li: From Multi-View to Hollow-3D: Hallucinated Hollow-3D R-CNN for 3D Object Detection. IEEE Transactions on Circuits and Systems for Video Technology 2021.
92	Test_dif		code	72.47 %	85.59 %	67.50 %	0.01 s	1 core @ 2.5 Ghz (C/C++)

93	SFA-GCL(80)		code	72.46 %	87.24 %	65.43 %	0.04 s	1 core @ 2.5 Ghz (C/C++)

94	SFA-GCL(80, k=4)		code	72.34 %	86.83 %	63.31 %	0.04 s	1 core @ 2.5 Ghz (C/C++)

95	SFA-GCL		code	72.34 %	87.19 %	65.26 %	0.04 s	1 core @ 2.5 Ghz (C/C++)

96	HA-PillarNet			72.27 %	82.02 %	67.48 %	0.05 s	1 core @ 2.5 Ghz (Python)

97	SeSame-point		code	72.25 %	84.16 %	65.39 %	N/A s	TITAN RTX @ 1.35 Ghz (Python)
H. O, C. Yang and K. Huh: SeSame: Simple, Easy 3D Object Detection with Point-Wise Semantics. Proceedings of the Asian Conference on Computer Vision (ACCV) 2024.
98	MonoPSR		code	72.08 %	82.06 %	62.43 %	0.2 s	GPU @ 3.5 Ghz (Python)
J. Ku, A. Pon and S. Waslander: Monocular 3D Object Detection Leveraging Accurate Proposals and Shape Reconstruction. CVPR 2019.
99	HMNet			72.00 %	85.29 %	66.07 %	0.1 s	1 core @ 2.5 Ghz (C/C++)

100	ARPNET			71.95 %	84.96 %	65.21 %	0.08 s	GPU @ 2.5 Ghz (Python + C/C++)
Y. Ye, C. Zhang and X. Hao: ARPNET: attention region proposal network for 3D object detection. Science China Information Sciences 2019.
101	SubCNN			71.72 %	79.36 %	62.74 %	2 s	GPU @ 3.5 Ghz (Python + C/C++)
Y. Xiang, W. Choi, Y. Lin and S. Savarese: Subcategory-aware Convolutional Neural Networks for Object Proposals and Detection. IEEE Winter Conference on Applications of Computer Vision (WACV) 2017.
102	STD		code	71.63 %	83.99 %	64.92 %	0.08 s	GPU @ 2.5 Ghz (Python + C/C++)
Z. Yang, Y. Sun, S. Liu, X. Shen and J. Jia: STD: Sparse-to-Dense 3D Object Detector for Point Cloud. ICCV 2019.
103	Anonymous			71.13 %	82.66 %	65.88 %	0.02 s	GPU @ 2.5 Ghz (Python)

104	R2Pfusion-Det			71.01 %	82.87 %	63.90 %	0.3 s	1 core @ 2.5 Ghz (C/C++)

105	IA-SSD (multi)		code	70.46 %	84.98 %	65.55 %	0.014 s	1 core @ 2.5 Ghz (C/C++)
Y. Zhang, Q. Hu, G. Xu, Y. Ma, J. Wan and Y. Guo: Not All Points Are Equal: Learning Highly Efficient Point-based Detectors for 3D LiDAR Point Clouds. CVPR 2022.
106	MGAF-3DSSD		code	70.41 %	86.42 %	63.26 %	0.1 s	1 core @ 2.5 Ghz (C/C++)
J. Li, H. Dai, L. Shao and Y. Ding: Anchor-free 3D Single Stage Detector with Mask-Guided Attention for Point Cloud. MM '21: The 29th ACM International Conference on Multimedia (ACM MM) 2021.
107	PCNet3D++			70.28 %	82.71 %	64.10 %	0.5 s	GPU @ 3.5 Ghz (Python)

108	AB3DMOT		code	70.18 %	82.86 %	63.55 %	0.0047s	1 core @ 2.5 Ghz (Python)
X. Weng and K. Kitani: A Baseline for 3D Multi-Object Tracking. arXiv:1907.03961 2019.
109	SFA-GCL(baseline)		code	69.91 %	86.64 %	60.86 %	0.04 s	1 core @ 2.5 Ghz (C/C++)

110	DensePointPillars			69.65 %	81.88 %	63.61 %	0.03 s	GPU @ 2.5 Ghz (Python)

111	SFA-GCL_dataaug		code	69.65 %	86.23 %	62.64 %	0.04 s	1 core @ 2.5 Ghz (C/C++)

112	SeSame-voxel		code	69.59 %	87.27 %	62.84 %	N/A s	TITAN RTX @ 1.35 Ghz (Python)
H. O, C. Yang and K. Huh: SeSame: Simple, Easy 3D Object Detection with Point-Wise Semantics. Proceedings of the Asian Conference on Computer Vision (ACCV) 2024.
113	SFA-GCL		code	69.58 %	86.24 %	62.53 %	0.04 s	1 core @ 2.5 Ghz (C/C++)

114	T-SSD			69.08 %	83.96 %	64.75 %	0.04	1 core @ 2.0 Ghz (C/C++)

115	PointPillars		code	68.98 %	83.97 %	62.17 %	16 ms	1080ti GPU and Intel i7 CPU
A. Lang, S. Vora, H. Caesar, L. Zhou, J. Yang and O. Beijbom: PointPillars: Fast Encoders for Object Detection from Point Clouds. CVPR 2019.
116	Vote3Deep			68.82 %	78.41 %	62.50 %	1.5 s	4 cores @ 2.5 Ghz (C/C++)
M. Engelcke, D. Rao, D. Zeng Wang, C. Hay Tong and I. Posner: Vote3Deep: Fast Object Detection in 3D Point Clouds Using Efficient Convolutional Neural Networks. ArXiv e-prints 2016.
117	geo-pillars			68.79 %	81.21 %	62.26 %	0.05 s	1 core @ 2.5 Ghz (C/C++)
ERROR: Wrong syntax in BIBTEX file.
118	3DOP		code	68.71 %	80.52 %	61.07 %	3s	GPU @ 2.5 Ghz (Matlab + C/C++)
X. Chen, K. Kundu, Y. Zhu, A. Berneshawi, H. Ma, S. Fidler and R. Urtasun: 3D Object Proposals for Accurate Object Class Detection. NIPS 2015.
119	Pose-RCNN			68.40 %	81.53 %	59.43 %	2 s	>8 cores @ 2.5 Ghz (Python)
M. Braun, Q. Rao, Y. Wang and F. Flohr: Pose-RCNN: Joint object detection and pose estimation using 3D object proposals. Intelligent Transportation Systems (ITSC), 2016 IEEE 19th International Conference on 2016.
120	PCNet3D			68.37 %	82.24 %	62.15 %	0.05 s	GPU @ 3.5 Ghz (Python)

121	EPNet++			68.30 %	80.27 %	63.00 %	0.1 s	GPU @ 2.5 Ghz (Python)
Z. Liu, T. Huang, B. Li, X. Chen, X. Wang and X. Bai: EPNet++: Cascade Bi-Directional Fusion for Multi-Modal 3D Object Detection. IEEE Transactions on Pattern Analysis and Machine Intelligence 2022.
122	TANet		code	68.20 %	82.24 %	62.13 %	0.035s	GPU @ 2.5 Ghz (Python + C/C++)
Z. Liu, X. Zhao, T. Huang, R. Hu, Y. Zhou and X. Bai: TANet: Robust 3D Object Detection from Point Clouds with Triple Attention. AAAI 2020.
123	AEPF			67.72 %	79.94 %	61.91 %	0.05 s	GPU @ 2.5 Ghz (Python)

124	IVA		code	67.57 %	78.48 %	58.83 %	0.4 s	GPU @ 2.5 Ghz (C/C++)
Y. Zhu, J. Wang, C. Zhao, H. Guo and H. Lu: Scale-adaptive Deconvolutional Regression Network for Pedestrian Detection. ACCV 2016. S. Ren, K. He, R. Girshick and J. Sun: Faster R-CNN: Towards real-time object detection with region proposal networks. Advances in neural information processing systems 2015.
125	Fade-kd			67.44 %	81.35 %	61.32 %	0.01 s	1 core @ 2.5 Ghz (C/C++)

126	DeepStereoOP			67.22 %	79.35 %	58.60 %	3.4 s	GPU @ 3.5 Ghz (Matlab + C/C++)
C. Pham and J. Jeon: Robust Object Proposals Re-ranking for Object Detection in Autonomous Driving Using Convolutional Neural Networks. Signal Processing: Image Communiation 2017.
127	Cube R-CNN		code	66.98 %	81.99 %	58.56 %	0.05 s	GPU @ 2.5 Ghz (Python)
G. Brazil, A. Kumar, J. Straub, N. Ravi, J. Johnson and G. Gkioxari: Omni3D: A Large Benchmark and Model for 3D Object Detection in the Wild. CVPR 2023.
128	SeSame-pillar		code	66.76 %	77.99 %	60.45 %	N/A s	TITAN RTX @ 1.35 Ghz (Python)
H. O, C. Yang and K. Huh: SeSame: Simple, Easy 3D Object Detection with Point-Wise Semantics. Proceedings of the Asian Conference on Computer Vision (ACCV) 2024.
129	FII-CenterNet		code	66.54 %	79.04 %	57.76 %	0.09 s	GPU @ 2.5 Ghz (Python)
S. Fan, F. Zhu, S. Chen, H. Zhang, B. Tian, Y. Lv and F. Wang: FII-CenterNet: An Anchor-Free Detector With Foreground Attention for Traffic Object Detection. IEEE Transactions on Vehicular Technology 2021.
130	epBRM		code	66.51 %	79.65 %	60.31 %	0.10 s	1 core @ 2.5 Ghz (C/C++)
K. Shin: Improving a Quality of 3D Object Detection by Spatial Transformation Mechanism. arXiv preprint arXiv:1910.04853 2019.
131	PFF3D		code	66.25 %	79.44 %	60.11 %	0.05 s	GPU @ 3.0 Ghz (Python + C/C++)
L. Wen and K. Jo: Fast and Accurate 3D Object Detection for Lidar-Camera-Based Autonomous Vehicles Using One Shared Voxel-Based Backbone. IEEE Access 2021.
132	DD3D		code	65.98 %	81.13 %	58.86 %	n/a s	1 core @ 2.5 Ghz (C/C++)
D. Park, R. Ambrus, V. Guizilini, J. Li and A. Gaidon: Is Pseudo-Lidar needed for Monocular 3D Object detection?. IEEE/CVF International Conference on Computer Vision (ICCV) .
133	PointRGBNet			65.98 %	79.87 %	59.75 %	0.08 s	4 cores @ 2.5 Ghz (Python + C/C++)
P. Xie Desheng: Real-time Detection of 3D Objects Based on Multi-Sensor Information Fusion. Automotive Engineering 2022.
134	BirdNet+		code	65.40 %	72.96 %	60.23 %	0.11 s	Titan Xp (PyTorch)
A. Barrera, J. Beltrán, C. Guindel, J. Iglesias and F. García: BirdNet+: Two-Stage 3D Object Detection in LiDAR through a Sparsity-Invariant Bird’s Eye View. IEEE Access 2021.
135	PillarHist			65.21 %	76.69 %	58.80 %	0.01 s	1 core @ 2.5 Ghz (C/C++)

136	Mono3D		code	65.15 %	77.19 %	57.88 %	4.2 s	GPU @ 2.5 Ghz (Matlab + C/C++)
X. Chen, K. Kundu, Z. Zhang, H. Ma, S. Fidler and R. Urtasun: Monocular 3D Object Detection for Autonomous Driving. CVPR 2016.
137	DMF			63.39 %	74.69 %	56.96 %	0.2 s	1 core @ 2.5 Ghz (Python + C/C++)
X. J. Chen and W. Xu: Disparity-Based Multiscale Fusion Network for Transportation Detection. IEEE Transactions on Intelligent Transportation Systems 2022.
138	PiFeNet		code	63.34 %	78.05 %	56.46 %	0.03 s	1 core @ 2.5 Ghz (C/C++)
D. Le, H. Shi, H. Rezatofighi and J. Cai: Accurate and Real-time 3D Pedestrian Detection Using an Efficient Attentive Pillar Network. IEEE Robotics and Automation Letters 2022.
139	Faster R-CNN		code	62.86 %	72.40 %	54.97 %	2 s	GPU @ 3.5 Ghz (Python + C/C++)
S. Ren, K. He, R. Girshick and J. Sun: Faster R-CNN: Towards Real- Time Object Detection with Region Proposal Networks. NIPS 2015.
140	SCNet			62.50 %	78.48 %	56.34 %	0.04 s	GPU @ 3.0 Ghz (Python)
Z. Wang, H. Fu, L. Wang, L. Xiao and B. Dai: SCNet: Subdivision Coding Network for Object Detection Based on 3D Point Cloud. IEEE Access 2019.
141	DSGN++		code	62.10 %	77.71 %	55.78 %	0.2 s	GeForce RTX 2080Ti
Y. Chen, S. Huang, S. Liu, B. Yu and J. Jia: DSGN++: Exploiting Visual-Spatial Relation for Stereo-Based 3D Detectors. IEEE Transactions on Pattern Analysis and Machine Intelligence 2022.
142	StereoDistill			61.46 %	80.92 %	54.64 %	0.4 s	1 core @ 2.5 Ghz (Python)
Z. Liu, X. Ye, X. Tan, D. Errui, Y. Zhou and X. Bai: StereoDistill: Pick the Cream from LiDAR for Distilling Stereo-based 3D Object Detection. Proceedings of the AAAI Conference on Artificial Intelligence 2023.
143	Fade			61.19 %	76.76 %	55.78 %	0.01 s	1 core @ 2.5 Ghz (C/C++)

144	AVOD-FPN		code	60.79 %	70.38 %	55.37 %	0.1 s	Titan X (Pascal)
J. Ku, M. Mozifian, J. Lee, A. Harakeh and S. Waslander: Joint 3D Proposal Generation and Object Detection from View Aggregation. IROS 2018.
145	SDP+CRC (ft)			60.72 %	75.63 %	53.00 %	0.6 s	GPU @ 2.5 Ghz (C/C++)
F. Yang, W. Choi and Y. Lin: Exploit All the Layers: Fast and Accurate CNN Object Detector with Scale Dependent Pooling and Cascaded Rejection Classifiers. Proceedings of the IEEE International Conference on Computer Vision and Pattern Recognition 2016.
146	PL++: PV-RCNN++			60.35 %	74.52 %	53.83 %	0.342 s	RTX 4060Ti (Python)

147	Complexer-YOLO			59.78 %	66.94 %	55.63 %	0.06 s	GPU @ 3.5 Ghz (C/C++)
M. Simon, K. Amende, A. Kraus, J. Honer, T. Samann, H. Kaulbersch, S. Milz and H. Michael Gross: Complexer-YOLO: Real-Time 3D Object Detection and Tracking on Semantic Point Clouds. The IEEE Conference on Computer Vision and Pattern Recognition (CVPR) Workshops 2019.
148	Mix-Teaching		code	58.65 %	75.15 %	50.54 %	30 s	1 core @ 2.5 Ghz (C/C++)
L. Yang, X. Zhang, L. Wang, M. Zhu, C. Zhang and J. Li: Mix-Teaching: A Simple, Unified and Effective Semi-Supervised Learning Framework for Monocular 3D Object Detection. ArXiv 2022.
149	Regionlets			58.52 %	71.12 %	50.83 %	1 s	>8 cores @ 2.5 Ghz (C/C++)
X. Wang, M. Yang, S. Zhu and Y. Lin: Regionlets for Generic Object Detection. T-PAMI 2015. W. Zou, X. Wang, M. Sun and Y. Lin: Generic Object Detection with Dense Neural Patterns and Regionlets. British Machine Vision Conference 2014. C. Long, X. Wang, G. Hua, M. Yang and Y. Lin: Accurate Object Detection with Location Relaxation and Regionlets Relocalization. Asian Conference on Computer Vision 2014.
150	MonoLiG		code	58.35 %	80.41 %	51.21 %	0.03 s	1 core @ 2.5 Ghz (C/C++)
A. Hekimoglu, M. Schmidt and A. Ramiro: Monocular 3D Object Detection with LiDAR Guided Semi Supervised Active Learning. 2023.
151	FRCNN+Or		code	57.01 %	70.99 %	50.14 %	0.09 s	Titan Xp GPU
C. Guindel, D. Martin and J. Armingol: Fast Joint Object Detection and Viewpoint Estimation for Traffic Scene Understanding. IEEE Intelligent Transportation Systems Magazine 2018. C. Guindel, D. Martin and J. Armingol: Joint Object Detection and Viewpoint Estimation using CNN features. IEEE International Conference on Vehicular Electronics and Safety (ICVES) 2017.
152	QD-3DT		code	56.51 %	75.55 %	49.70 %	0.03 s	GPU @ 2.5 Ghz (Python)
H. Hu, Y. Yang, T. Fischer, F. Yu, T. Darrell and M. Sun: Monocular Quasi-Dense 3D Object Tracking. ArXiv:2103.07351 2021.
153	MonoPair			56.37 %	74.77 %	48.37 %	0.06 s	GPU @ 2.5 Ghz (Python + C/C++)
Y. Chen, L. Tai, K. Sun and M. Li: MonoPair: Monocular 3D Object Detection Using Pairwise Spatial Relationships. The IEEE Conference on Computer Vision and Pattern Recognition (CVPR) 2020.
154	MLOD		code	56.04 %	75.35 %	49.11 %	0.12 s	GPU @ 1.5 Ghz (Python)
J. Deng and K. Czarnecki: MLOD: A multi-view 3D object detection based on robust feature fusion method. arXiv preprint arXiv:1909.04163 2019.
155	SeSame-voxel w/score		code	55.27 %	67.26 %	50.21 %	N/A s	GPU @ 1.5 Ghz (Python)
H. O, C. Yang and K. Huh: SeSame: Simple, Easy 3D Object Detection with Point-Wise Semantics. Proceedings of the Asian Conference on Computer Vision (ACCV) 2024.
156	MonoAFKD			54.77 %	74.42 %	48.08 %	0.03 s	1 core @ 2.5 Ghz (C/C++)

157	MonoFlex			54.76 %	72.41 %	46.21 %	0.03 s	GPU @ 2.5 Ghz (Python)
Y. Zhang, J. Lu and J. Zhou: Objects are Different: Flexible Monocular 3D Object Detection. CVPR 2021.
158	MonoLSS			54.63 %	74.54 %	47.98 %	0.04 s	1 core @ 2.5 Ghz (Python)
Z. Li, J. Jia and Y. Shi: MonoLSS: Learnable Sample Selection For Monocular 3D Detection. International Conference on 3D Vision 2024.
159	BirdNet+ (legacy)		code	54.61 %	74.97 %	50.29 %	0.1 s	Titan Xp (PyTorch)
A. Barrera, C. Guindel, J. Beltrán and F. García: BirdNet+: End-to-End 3D Object Detection in LiDAR Bird’s Eye View. 2020 IEEE 23rd International Conference on Intelligent Transportation Systems (ITSC) 2020.
160	AM			54.58 %	72.88 %	48.14 %	0.04 s	1 core @ 2.5 Ghz (C/C++)

161	MMLAB LIGA-Stereo		code	54.57 %	74.40 %	48.11 %	0.4 s	1 core @ 2.5 Ghz (Python + C/C++)
X. Guo, S. Shi, X. Wang and H. Li: LIGA-Stereo: Learning LiDAR Geometry Aware Representations for Stereo-based 3D Detector. Proceedings of the IEEE/CVF International Conference on Computer Vision (ICCV) 2021.
162	HomoLoss(monoflex)		code	54.12 %	70.14 %	46.16 %	0.04 s	1 core @ 2.5 Ghz (Python)
J. Gu, B. Wu, L. Fan, J. Huang, S. Cao, Z. Xiang and X. Hua: Homography Loss for Monocular 3D Object Detection. Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) 2022.
163	MonoUNI		code	53.71 %	71.68 %	45.26 %	0.04 s	1 core @ 2.5 Ghz (Python)
J. Jia, Z. Li and Y. Shi: MonoUNI: A Unified Vehicle and Infrastructure-side Monocular 3D Object Detection Network with Sufficient Depth Clues. Thirty-seventh Conference on Neural Information Processing Systems 2023.
164	monodle		code	53.29 %	70.78 %	45.01 %	0.04 s	GPU @ 2.5 Ghz (Python)
X. Ma, Y. Zhang, D. Xu, D. Zhou, S. Yi, H. Li and W. Ouyang: Delving into Localization Errors for Monocular 3D Object Detection. CVPR 2021 .
165	LPCG-Monoflex		code	53.04 %	72.36 %	46.11 %	0.03 s	1 core @ 2.5 Ghz (C/C++)
L. Peng, F. Liu, Z. Yu, S. Yan, D. Deng, Z. Yang, H. Liu and D. Cai: Lidar Point Cloud Guided Monocular 3D Object Detection. ECCV 2022.
166	AVOD		code	52.60 %	66.45 %	46.39 %	0.08 s	Titan X (pascal)
J. Ku, M. Mozifian, J. Lee, A. Harakeh and S. Waslander: Joint 3D Proposal Generation and Object Detection from View Aggregation. IROS 2018.
167	CMKD		code	51.76 %	73.18 %	45.37 %	0.1 s	1 core @ 2.5 Ghz (C/C++)
Y. Hong, H. Dai and Y. Ding: Cross-Modality Knowledge Distillation Network for Monocular 3D Object Detection. ECCV 2022.
168	MonOAPC			51.68 %	68.18 %	44.08 %	0035 s	1 core @ 2.5 Ghz (Python)
H. Yao, J. Chen, Z. Wang, X. Wang, P. Han, X. Chai and Y. Qiu: Occlusion-Aware Plane-Constraints for Monocular 3D Object Detection. IEEE Transactions on Intelligent Transportation Systems 2023.
169	UniCuboid			51.59 %	73.71 %	44.86 %	0.1 s	1 core @ 2.5 Ghz (C/C++)

170	MonoDDE			51.10 %	70.85 %	44.02 %	0.04 s	1 core @ 2.5 Ghz (Python)
Z. Li, Z. Qu, Y. Zhou, J. Liu, H. Wang and L. Jiang: Diversity Matters: Fully Exploiting Depth Clues for Reliable Monocular 3D Object Detection. CVPR 2022.
171	MonoDTR			49.48 %	64.93 %	42.76 %	0.04 s	1 core @ 2.5 Ghz (C/C++)
K. Huang, T. Wu, H. Su and W. Hsu: MonoDTR: Monocular 3D Object Detection with Depth-Aware Transformer. CVPR 2022.
172	MonoRUn		code	49.13 %	67.47 %	43.41 %	0.07 s	GPU @ 2.5 Ghz (Python + C/C++)
H. Chen, Y. Huang, W. Tian, Z. Gao and L. Xiong: MonoRUn: Monocular 3D Object Detection by Reconstruction and Uncertainty Propagation. IEEE Conference on Computer Vision and Pattern Recognition (CVPR) 2021.
173	MonoRCNN++		code	48.84 %	67.78 %	42.44 %	0.07 s	GPU @ 2.5 Ghz (Python)
X. Shi, Z. Chen and T. Kim: Multivariate Probabilistic Monocular 3D Object Detection. WACV 2023.
174	CG-Stereo			48.46 %	69.98 %	42.41 %	0.57 s	GeForce RTX 2080 Ti
C. Li, J. Ku and S. Waslander: Confidence Guided Stereo 3D Object Detection with Split Depth Estimation. IROS 2020.
175	BirdNet			47.64 %	64.91 %	44.59 %	0.11 s	Titan Xp (Caffe)
J. Beltrán, C. Guindel, F. Moreno, D. Cruzado, F. García and A. Escalera: BirdNet: A 3D Object Detection Framework from LiDAR Information. 2018 21st International Conference on Intelligent Transportation Systems (ITSC) 2018.
176	monospb			47.40 %	68.42 %	41.60 %	0.01 s	1 core @ 2.5 Ghz (Python)

177	DEVIANT		code	46.42 %	67.71 %	39.44 %	0.04 s	1 GPU (Python)
A. Kumar, G. Brazil, E. Corona, A. Parchami and X. Liu: DEVIANT: Depth EquiVarIAnt NeTwork for Monocular 3D Object Detection. European Conference on Computer Vision (ECCV) 2022.
178	Disp R-CNN (velo)		code	46.37 %	63.22 %	40.15 %	0.387 s	GPU @ 2.5 Ghz (Python + C/C++)
J. Sun, L. Chen, Y. Xie, S. Zhang, Q. Jiang, X. Zhou and H. Bao: Disp R-CNN: Stereo 3D Object Detection via Shape Prior Guided Instance Disparity Estimation. CVPR 2020.
179	Disp R-CNN		code	46.37 %	63.24 %	40.15 %	0.387 s	GPU @ 2.5 Ghz (Python + C/C++)
J. Sun, L. Chen, Y. Xie, S. Zhang, Q. Jiang, X. Zhou and H. Bao: Disp R-CNN: Stereo 3D Object Detection via Shape Prior Guided Instance Disparity Estimation. CVPR 2020.
180	SAKD-MR-Res18			45.72 %	64.63 %	39.60 %	0.03 s	1 core @ 2.5 Ghz (C/C++)

181	SparsePool		code	44.57 %	60.53 %	40.37 %	0.13 s	8 cores @ 2.5 Ghz (Python)
Z. Wang, W. Zhan and M. Tomizuka: Fusing bird view lidar point cloud and front view camera image for deep object detection. arXiv preprint arXiv:1711.06703 2017.
182	MonoCoP			44.17 %	64.73 %	39.45 %	0.01 s	1 core @ 2.5 Ghz (C/C++)

183	Shift R-CNN (mono)		code	42.96 %	63.24 %	38.22 %	0.25 s	GPU @ 1.5 Ghz (Python)
A. Naiden, V. Paunescu, G. Kim, B. Jeon and M. Leordeanu: Shift R-CNN: Deep Monocular 3D Object Detection With Closed-form Geometric Constraints. ICIP 2019.
184	D4LCN		code	42.86 %	65.29 %	36.29 %	0.2 s	GPU @ 2.5 Ghz (Python + C/C++)
M. Ding, Y. Huo, H. Yi, Z. Wang, J. Shi, Z. Lu and P. Luo: Learning Depth-Guided Convolutions for Monocular 3D Object Detection. CVPR 2020.
185	GUPNet		code	42.78 %	67.11 %	37.94 %	NA s	1 core @ 2.5 Ghz (Python + C/C++)
Y. Lu, X. Ma, L. Yang, T. Zhang, Y. Liu, Q. Chu, J. Yan and W. Ouyang: Geometry Uncertainty Projection Network for Monocular 3D Object Detection. arXiv preprint arXiv:2107.13774 2021.
186	CPD++(unsupervised)		code	42.56 %	57.51 %	35.46 %	0.1 s	GPU @ >3.5 Ghz (Python)

187	M3D-RPN		code	41.54 %	61.54 %	35.23 %	0.16 s	GPU @ 1.5 Ghz (Python)
G. Brazil and X. Liu: M3D-RPN: Monocular 3D Region Proposal Network for Object Detection . ICCV 2019 .
188	MonoEF			41.19 %	51.06 %	35.70 %	0.03 s	1 core @ 2.5 Ghz (Python)
Y. Zhou, Y. He, H. Zhu, C. Wang, H. Li and Q. Jiang: Monocular 3D Object Detection: An Extrinsic Parameter Free Approach. Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) 2021.
189	PS-fld		code	41.13 %	58.13 %	35.90 %	0.25 s	1 core @ 2.5 Ghz (C/C++)
Y. Chen, H. Dai and Y. Ding: Pseudo-Stereo for Monocular 3D Object Detection in Autonomous Driving. Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) 2022.
190	Plane-Constraints		code	41.01 %	58.71 %	35.35 %	0.05 s	4 cores @ 3.0 Ghz (Python)
H. Yao, J. Chen, Z. Wang, X. Wang, X. Chai, Y. Qiu and P. Han: Vertex points are not enough: Monocular 3D object detection via intra-and inter-plane constraints. Neural Networks 2023.
191	MV-RGBD-RF			40.94 %	51.10 %	34.83 %	4 s	4 cores @ 2.5 Ghz (C/C++)
A. Gonzalez, D. Vazquez, A. Lopez and J. Amores: On-Board Object Detection: Multicue, Multimodal, and Multiview Random Forest of Local Experts.. IEEE Trans. on Cybernetics 2016. A. Gonzalez, G. Villalonga, J. Xu, D. Vazquez, J. Amores and A. Lopez: Multiview Random Forest of Local Experts Combining RGB and LIDAR data for Pedestrian Detection. IEEE Intelligent Vehicles Symposium (IV) 2015.
192	MonoMH		code	39.69 %	59.04 %	34.04 %	0.04 s	1 core @ 2.5 Ghz (Python)

193	MonoFRD			38.98 %	55.86 %	34.32 %	0.01 s	1 core @ 2.5 Ghz (C/C++)
Z. Gong, Y. Zhao, F. Zhang, G. Gui, B. Chen, L. Yu, H. Wang, C. Yang and W. Gui: Color intuitive feature guided depth-height fusion and volume rendering for monocular 3D object detection. IEEE Transactions on Intelligent Vehicles(Major Revison) 2024.
194	DDMP-3D			38.62 %	58.70 %	34.10 %	0.18 s	1 core @ 2.5 Ghz (Python)
L. Wang, L. Du, X. Ye, Y. Fu, G. Guo, X. Xue, J. Feng and L. Zhang: Depth-conditioned Dynamic Message Propagation for Monocular 3D Object Detection. CVPR 2020.
195	CMAN			38.36 %	58.12 %	31.79 %	0.15 s	1 core @ 2.5 Ghz (Python)
C. Yuanzhouhan Cao: CMAN: Leaning Global Structure Correlation for Monocular 3D Object Detection. IEEE Trans. Intell. Transport. Syst. 2022.
196	OPA-3D		code	38.35 %	55.98 %	33.83 %	0.04 s	1 core @ 3.5 Ghz (Python)
Y. Su, Y. Di, G. Zhai, F. Manhardt, J. Rambach, B. Busam, D. Stricker and F. Tombari: OPA-3D: Occlusion-Aware Pixel-Wise Aggregation for Monocular 3D Object Detection. IEEE Robotics and Automation Letters 2023.
197	Aug3D-RPN			36.69 %	51.49 %	30.04 %	0.08 s	1 core @ 2.5 Ghz (C/C++)
C. He, J. Huang, X. Hua and L. Zhang: Aug3D-RPN: Improving Monocular 3D Object Detection by Synthetic Images with Virtual Depth. 2021.
198	SparsePool		code	36.26 %	44.21 %	32.57 %	0.13 s	8 cores @ 2.5 Ghz (Python)
Z. Wang, W. Zhan and M. Tomizuka: Fusing bird view lidar point cloud and front view camera image for deep object detection. arXiv preprint arXiv:1711.06703 2017.
199	SS3D			35.48 %	52.97 %	31.07 %	48 ms	Tesla V100 (Python)
E. Jörgensen, C. Zach and F. Kahl: Monocular 3D Object Detection and Box Fitting Trained End-to-End Using Intersection-over-Union Loss. CoRR 2019.
200	mdab			35.46 %	52.35 %	30.87 %	0.02 s	1 core @ 2.5 Ghz (Python)

201	mdab			35.46 %	52.35 %	30.87 %	22 s	1 core @ 2.5 Ghz (C/C++)

202	DSGN		code	35.15 %	49.10 %	31.41 %	0.67 s	NVIDIA Tesla V100
Y. Chen, S. Liu, X. Shen and J. Jia: DSGN: Deep Stereo Geometry Network for 3D Object Detection. CVPR 2020.
203	pAUCEnsT			34.90 %	50.51 %	30.35 %	60 s	1 core @ 2.5 Ghz (Matlab + C/C++)
S. Paisitkriangkrai, C. Shen and A. Hengel: Pedestrian Detection with Spatially Pooled Features and Structured Ensemble Learning. arXiv 2014.
204	TopNet-Retina			31.98 %	47.51 %	29.84 %	52ms	GeForce 1080Ti (tensorflow-gpu, v1.12)
S. Wirges, T. Fischer, C. Stiller and J. Frias: Object Detection and Classification in Occupancy Grid Maps Using Deep Convolutional Networks. 2018 21st International Conference on Intelligent Transportation Systems (ITSC) 2018.
205	DFR-Net			31.93 %	48.34 %	27.95 %	0.18 s	1080 Ti (Pytorch)
Z. Zou, X. Ye, L. Du, X. Cheng, X. Tan, L. Zhang, J. Feng, X. Xue and E. Ding: The devil is in the task: Exploiting reciprocal appearance-localization features for monocular 3d object detection . ICCV 2021.
206	AMNet+DDAD15M		code	31.01 %	45.93 %	27.06 %	0.03 s	1 core @ 2.5 Ghz (C/C++)
H. Pan, Y. Jia, J. Wang and W. Sun: MonoAMNet: Three-Stage Real-Time Monocular 3D Object Detection With Adaptive Methods. IEEE Transactions on Intelligent Transportation Systems 2025.
207	PS-SVDM			30.95 %	46.46 %	27.00 %	1 s	1 core @ 2.5 Ghz (Python)
Y. Shi: SVDM: Single-View Diffusion Model for Pseudo-Stereo 3D Object Detection. arXiv preprint arXiv:2307.02270 2023.
208	PS-SVDM			30.16 %	46.35 %	26.13 %	1 s	1 core @ 2.5 Ghz (Python)
Y. Shi: SVDM: Single-View Diffusion Model for Pseudo-Stereo 3D Object Detection. arXiv preprint arXiv:2307.02270 2023.
209	CIE			30.10 %	38.03 %	26.99 %	0.1 s	1 core @ 2.5 Ghz (C/C++)
Anonymities: Consistency of Implicit and Explicit Features Matters for Monocular 3D Object Detection. arXiv preprint arXiv:2207.07933 2022.
210	MonoNeRD		code	29.89 %	45.35 %	26.49 %	na s	1 core @ 2.5 Ghz (Python)
J. Xu, L. Peng, H. Cheng, H. Li, W. Qian, K. Li, W. Wang and D. Cai: MonoNeRD: NeRF-like Representations for Monocular 3D Object Detection. ICCV 2023.
211	AMNet		code	29.62 %	43.82 %	25.55 %	0.03 s	GPU @ 1.0 Ghz (Python)
H. Pan, Y. Jia, J. Wang and W. Sun: MonoAMNet: Three-Stage Real-Time Monocular 3D Object Detection With Adaptive Methods. IEEE Transactions on Intelligent Transportation Systems 2025.
212	OC Stereo		code	28.76 %	43.18 %	24.80 %	0.35 s	1 core @ 2.5 Ghz (Python + C/C++)
A. Pon, J. Ku, C. Li and S. Waslander: Object-Centric Stereo Matching for 3D Object Detection. ICRA 2020.
213	Vote3D			27.99 %	39.81 %	25.19 %	0.5 s	4 cores @ 2.8 Ghz (C/C++)
D. Wang and I. Posner: Voting for Voting in Online Point Cloud Object Detection. Proceedings of Robotics: Science and Systems 2015.
214	SGM3D		code	27.89 %	42.21 %	24.73 %	0.03 s	1 core @ 2.5 Ghz (C/C++)
Z. Zhou, L. Du, X. Ye, Z. Zou, X. Tan, L. Zhang, X. Xue and J. Feng: SGM3D: Stereo Guided Monocular 3D Object Detection. RA-L 2022.
215	LSVM-MDPM-us		code	27.81 %	37.66 %	24.83 %	10 s	4 cores @ 3.0 Ghz (C/C++)
P. Felzenszwalb, R. Girshick, D. McAllester and D. Ramanan: Object Detection with Discriminatively Trained Part-Based Models. PAMI 2010.
216	DPM-VOC+VP			27.73 %	41.58 %	24.61 %	8 s	1 core @ 2.5 Ghz (C/C++)
B. Pepik, M. Stark, P. Gehler and B. Schiele: Multi-view and 3D Deformable Part Models. IEEE Transactions on Pattern Analysis and Machine Intelligence (TPAMI) 2015.
217	RefinedMPL			27.17 %	44.47 %	22.84 %	0.15 s	GPU @ 2.5 Ghz (Python + C/C++)
J. Vianney, S. Aich and B. Liu: RefinedMPL: Refined Monocular PseudoLiDAR for 3D Object Detection in Autonomous Driving. arXiv preprint arXiv:1911.09712 2019.
218	CaDDN		code	27.13 %	40.03 %	23.23 %	0.63 s	GPU @ 2.5 Ghz (Python)
C. Reading, A. Harakeh, J. Chae and S. Waslander: Categorical Depth Distribution Network for Monocular 3D Object Detection. CVPR 2021.
219	PGD-FCOS3D		code	26.48 %	44.28 %	23.03 %	0.03 s	1 core @ 2.5 Ghz (C/C++)
T. Wang, X. Zhu, J. Pang and D. Lin: Probabilistic and Geometric Depth: Detecting Objects in Perspective. Conference on Robot Learning (CoRL) 2021.
220	SeSame-pillar w/scor		code	26.32 %	23.56 %	24.54 %	N/A s	1 core @ 2.5 Ghz (C/C++)
H. O, C. Yang and K. Huh: SeSame: Simple, Easy 3D Object Detection with Point-Wise Semantics. Proceedings of the Asian Conference on Computer Vision (ACCV) 2024.
221	LSVM-MDPM-sv			26.05 %	35.70 %	23.56 %	10 s	4 cores @ 3.0 Ghz (C/C++)
P. Felzenszwalb, R. Girshick, D. McAllester and D. Ramanan: Object Detection with Discriminatively Trained Part-Based Models. PAMI 2010. A. Geiger, C. Wojek and R. Urtasun: Joint 3D Estimation of Objects and Scene Layout. NIPS 2011.
222	DPM-C8B1			25.57 %	41.47 %	21.93 %	15 s	4 cores @ 2.5 Ghz (Matlab + C/C++)
J. Yebes, L. Bergasa and M. García-Garrido: Visual Object Recognition with 3D-Aware Features in KITTI Urban Scenes. Sensors 2015. J. Yebes, L. Bergasa, R. Arroyo and A. Lázaro: Supervised learning and evaluation of KITTI's cars detector with DPM. IV 2014.
223	LLW			25.17 %	38.19 %	22.09 %	0.1 s	1 core @ 2.5 Ghz (C/C++)

224	FMF-occlusion-net			23.59 %	37.41 %	21.20 %	0.16 s	1 core @ 2.5 Ghz (Python + C/C++)
H. Liu, H. Liu, Y. Wang, F. Sun and W. Huang: Fine-grained Multi-level Fusion for Anti- occlusion Monocular 3D Object Detection. IEEE Transactions on Image Processing 2022.
225	RT3D-GMP			22.90 %	33.64 %	19.87 %	0.06 s	GPU @ 2.5 Ghz (Python + C/C++)
H. Königshof and C. Stiller: Learning-Based Shape Estimation with Grid Map Patches for Realtime 3D Object Detection for Automated Driving. 2020 IEEE 23rd International Conference on Intelligent Transportation Systems (ITSC) 2020.
226	mBoW			17.63 %	26.66 %	16.02 %	10 s	1 core @ 2.5 Ghz (C/C++)
J. Behley, V. Steinhage and A. Cremers: Laser-based Segment Classification Using a Mixture of Bag-of-Words. Proc. of the IEEE/RSJ International Conference on Intelligent Robots and Systems (IROS) 2013.
227	MDSNet			16.64 %	28.23 %	14.14 %	0.05 s	1 core @ 2.5 Ghz (Python)
Z. Xie, Y. Song, J. Wu, Z. Li, C. Song and Z. Xu: MDS-Net: Multi-Scale Depth Stratification 3D Object Detection from Monocular Images. Sensors 2022.
228	TopNet-HighRes			13.98 %	22.86 %	14.52 %	101ms	NVIDIA GeForce 1080 Ti (tensorflow-gpu)
S. Wirges, T. Fischer, C. Stiller and J. Frias: Object Detection and Classification in Occupancy Grid Maps Using Deep Convolutional Networks. 2018 21st International Conference on Intelligent Transportation Systems (ITSC) 2018.
229	ESGN			13.45 %	21.13 %	11.72 %	0.06 s	GPU @ 2.5 Ghz (Python + C/C++)
A. Gao, Y. Pang, J. Nie, Z. Shao, J. Cao, Y. Guo and X. Li: ESGN: Efficient Stereo Geometry Network for Fast 3D Object Detection. IEEE Transactions on Circuits and Systems for Video Technology 2022.
230	RT3DStereo			12.96 %	19.58 %	11.47 %	0.08 s	GPU @ 2.5 Ghz (C/C++)
H. Königshof, N. Salscheider and C. Stiller: Realtime 3D Object Detection for Automated Driving Using Stereo Vision and Semantic Information. Proc. IEEE Intl. Conf. Intelligent Transportation Systems 2019.
231	TopNet-UncEst			12.00 %	18.14 %	11.85 %	0.09 s	NVIDIA GeForce 1080 Ti (tensorflow-gpu)
S. Wirges, M. Braun, M. Lauer and C. Stiller: Capturing Object Detection Uncertainty in Multi-Layer Grid Maps. 2019.
232	SeSame-point w/score		code	10.20 %	12.43 %	9.34 %	N/A s	1 core @ 1.5 Ghz (Python)
H. O, C. Yang and K. Huh: SeSame: Simple, Easy 3D Object Detection with Point-Wise Semantics. Proceedings of the Asian Conference on Computer Vision (ACCV) 2024.
233	CPD(unsupervised)		code	9.81 %	13.96 %	8.76 %	0.1 s	GPU @ >3.5 Ghz (Python + C/C++)
H. Wu, S. Zhao, X. Huang, C. Wen, X. Li and C. Wang: Commonsense Prototype for Outdoor Unsupervised 3D Object Detection. CVPR 2024.
234	GATE3D		code	0.68 %	0.00 %	0.83 %	0.01 s	1 core @ 2.5 Ghz (C/C++)

235	YOLOv2		code	0.06 %	0.15 %	0.07 %	0.02 s	GPU @ 3.5 Ghz (C/C++)
J. Redmon, S. Divvala, R. Girshick and A. Farhadi: You only look once: Unified, real-time object detection. Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition 2016. J. Redmon and A. Farhadi: YOLO9000: Better, Faster, Stronger. Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition 2017.
236	TopNet-DecayRate			0.04 %	0.00 %	0.04 %	92 ms	NVIDIA GeForce 1080 Ti (tensorflow-gpu)
S. Wirges, T. Fischer, C. Stiller and J. Frias: Object Detection and Classification in Occupancy Grid Maps Using Deep Convolutional Networks. 2018 21st International Conference on Intelligent Transportation Systems (ITSC) 2018.
237	f3sd		code	0.01 %	0.02 %	0.02 %	1.67 s	1 core @ 2.5 Ghz (C/C++)

238	DA3D+KM3D+v2-99		code	0.00 %	0.00 %	0.00 %	0.120s	GPU @ 2.5 Ghz (Python)
Y. Jia, J. Wang, H. Pan and W. Sun: Enhancing Monocular 3-D Object Detection Through Data Augmentation Strategies. IEEE Transactions on Instrumentation and Measurement 2024.
239	DA3D+KM3D		code	0.00 %	0.00 %	0.00 %	0.02 s	GPU @ 2.5 Ghz (Python)
Y. Jia, J. Wang, H. Pan and W. Sun: Enhancing Monocular 3-D Object Detection Through Data Augmentation Strategies. IEEE Transactions on Instrumentation and Measurement 2024.
240	DA3D		code	0.00 %	0.00 %	0.00 %	0.03 s	1 core @ 2.5 Ghz (Python)
Y. Jia, J. Wang, H. Pan and W. Sun: Enhancing Monocular 3-D Object Detection Through Data Augmentation Strategies. IEEE Transactions on Instrumentation and Measurement 2024.

	Method	Setting	Code	Moderate	Easy	Hard	Runtime	Environment
1	ViKIENet			97.90 %	98.59 %	92.98 %	0.04 s	1 core @ 2.5 Ghz (C/C++)

2	LVP			97.66 %	98.68 %	92.81 %	0.04 s	1 core @ 2.5 Ghz (C/C++)
Y. Chen, G. Cai, Z. Song, Z. Liu, B. Zeng, J. Li and Z. Wang: LVP: Leverage Virtual Points in Multi- modal Early Fusion for 3D Object Detection. IEEE Transactions on Geoscience and Remote Sensing 2024.
3	SCEMF			97.45 %	98.61 %	94.40 %	1 s	1 core @ 2.5 Ghz (C/C++)

4	OGMMDet		code	97.43 %	98.44 %	92.72 %	0.01 s	1 core @ 2.5 Ghz (C/C++)

5	UDeerPEP		code	97.39 %	98.40 %	94.80 %	0.1 s	1 core @ 2.5 Ghz (C/C++)
Z. Dong, H. Ji, X. Huang, W. Zhang, X. Zhan and J. Chen: PeP: a Point enhanced Painting method for unified point cloud tasks. 2023.
6	RM3D			97.16 %	95.91 %	92.38 %	0.18 s	1 core @ 2.5 Ghz (C/C++)

7	ANM		code	97.12 %	97.65 %	94.40 %	ANM	ANM

8	ViKIENet-R			97.08 %	95.78 %	92.11 %	0.06 s	1 core @ 2.5 Ghz (C/C++)

9	MuStD			97.03 %	97.88 %	93.74 %	67 ms	>8 cores @ 2.5 Ghz (Python)

10	VirConv-S		code	96.46 %	96.99 %	93.74 %	0.09 s	1 core @ 2.5 Ghz (C/C++)
H. Wu, C. Wen, S. Shi and C. Wang: Virtual Sparse Convolution for Multimodal 3D Object Detection. CVPR 2023.
11	GraR-VoI		code	96.29 %	96.81 %	91.06 %	0.07 s	1 core @ 2.5 Ghz (Python + C/C++)
H. Yang, Z. Liu, X. Wu, W. Wang, W. Qian, X. He and D. Cai: Graph R-CNN: Towards Accurate 3D Object Detection with Semantic-Decorated Local Graph. ECCV 2022.
12	MLF-DET			96.09 %	96.87 %	88.78 %	0.09 s	1 core @ 2.5 Ghz (C/C++)
Z. Lin, Y. Shen, S. Zhou, S. Chen and N. Zheng: MLF-DET: Multi-Level Fusion for Cross- Modal 3D Object Detection. International Conference on Artificial Neural Networks 2023.
13	GraR-Po		code	96.09 %	96.83 %	90.99 %	0.06 s	1 core @ 2.5 Ghz (Python + C/C++)
H. Yang, Z. Liu, X. Wu, W. Wang, W. Qian, X. He and D. Cai: Graph R-CNN: Towards Accurate 3D Object Detection with Semantic-Decorated Local Graph. ECCV 2022.
14	LPFusion			96.09 %	96.71 %	93.15 %	0.1 s	1 core @ 2.5 Ghz (C/C++)

15	SFD		code	96.05 %	98.95 %	90.96 %	0.1 s	1 core @ 2.5 Ghz (Python + C/C++)
X. Wu, L. Peng, H. Yang, L. Xie, C. Huang, C. Deng, H. Liu and D. Cai: Sparse Fuse Dense: Towards High Quality 3D Detection with Depth Completion. CVPR 2022.
16	VPFNet		code	96.04 %	96.63 %	90.99 %	0.06 s	2 cores @ 2.5 Ghz (Python)
H. Zhu, J. Deng, Y. Zhang, J. Ji, Q. Mao, H. Li and Y. Zhang: VPFNet: Improving 3D Object Detection with Virtual Point based LiDAR and Stereo Data Fusion. IEEE Transactions on Multimedia 2022.
17	VirConv-T		code	96.01 %	98.64 %	93.12 %	0.09 s	1 core @ 2.5 Ghz (C/C++)
H. Wu, C. Wen, S. Shi and C. Wang: Virtual Sparse Convolution for Multimodal 3D Object Detection. CVPR 2023.
18	TED		code	95.96 %	96.63 %	93.24 %	0.1 s	1 core @ 2.5 Ghz (C/C++)
H. Wu, C. Wen, W. Li, R. Yang and C. Wang: Transformation-Equivariant 3D Object Detection for Autonomous Driving. AAAI 2023.
19	LPFusion_three_class			95.96 %	98.72 %	93.08 %	0.1 s	1 core @ 2.5 Ghz (C/C++)

20	RDIoU		code	95.95 %	98.77 %	90.90 %	0.03 s	1 core @ 2.5 Ghz (Python + C/C++)
H. Sheng, S. Cai, N. Zhao, B. Deng, J. Huang, X. Hua, M. Zhao and G. Lee: Rethinking IoU-based Optimization for Single- stage 3D Object Detection. ECCV 2022.
21	ACFNet			95.95 %	96.64 %	93.17 %	0.11 s	1 core @ 2.5 Ghz (C/C++)
Y. Tian, X. Zhang, X. Wang, J. Xu, J. Wang, R. Ai, W. Gu and W. Ding: ACF-Net: Asymmetric Cascade Fusion for 3D Detection With LiDAR Point Clouds and Images. IEEE Transactions on Intelligent Vehicles 2023.
22	CLOCs		code	95.93 %	96.77 %	90.93 %	0.1 s	1 core @ 2.5 Ghz (Python)
S. Pang, D. Morris and H. Radha: CLOCs: Camera-LiDAR Object Candidates Fusion for 3D Object Detection . 2020 IEEE/RSJ International Conference on Intelligent Robots and Systems (IROS) 2020.
23	GraR-Vo		code	95.92 %	96.66 %	92.78 %	0.04 s	1 core @ 2.5 Ghz (C/C++)
H. Yang, Z. Liu, X. Wu, W. Wang, W. Qian, X. He and D. Cai: Graph R-CNN: Towards Accurate 3D Object Detection with Semantic-Decorated Local Graph. ECCV 2022.
24	SFD++			95.90 %	98.90 %	93.04 %	0.06 s	1 core @ 2.5 Ghz (C/C++)

25	BFT3D			95.87 %	96.94 %	88.61 %	0.15 s	1 core @ 2.5 Ghz (C/C++)

26	LumiNet		code	95.87 %	99.09 %	88.47 %	0.1 s	1 core @ 2.5 Ghz (Python)

27	UPIDet		code	95.85 %	96.25 %	93.17 %	0.11 s	1 core @ 2.5 Ghz (C/C++)
Y. Zhang, Q. Zhang, J. Hou, Y. Yuan and G. Xing: Unleash the Potential of Image Branch for Cross-modal 3D Object Detection. Thirty-seventh Conference on Neural Information Processing Systems 2023.
28	PVT-SSD			95.83 %	96.74 %	90.58 %	0.05 s	1 core @ 2.5 Ghz (Python + C/C++)
H. Yang, W. Wang, M. Chen, B. Lin, T. He, H. Chen, X. He and W. Ouyang: PVT-SSD: Single-Stage 3D Object Detector with Point-Voxel Transformer. CVPR 2023.
29	CLOCs_PVCas		code	95.79 %	96.74 %	90.81 %	0.1 s	1 core @ 2.5 Ghz (Python)
S. Pang, D. Morris and H. Radha: CLOCs: Camera-LiDAR Object Candidates Fusion for 3D Object Detection . 2020 IEEE/RSJ International Conference on Intelligent Robots and Systems (IROS) 2020.
30	MPCF		code	95.78 %	98.94 %	90.88 %	0.08 s	1 core @ 2.5 Ghz (C/C++)
P. Gao and P. Zhang: MPCF: Multi-Phase Consolidated Fusion for Multi-Modal 3D Object Detection with Pseudo Point Cloud. 2024.
31	DDF			95.78 %	96.52 %	93.17 %	0.1 s	1 core @ 2.5 Ghz (Python)

32	3D Dual-Fusion		code	95.76 %	96.53 %	93.01 %	0.1 s	1 core @ 2.5 Ghz (Python)
Y. Kim, K. Park, M. Kim, D. Kum and J. Choi: 3D Dual-Fusion: Dual-Domain Dual-Query Camera-LiDAR Fusion for 3D Object Detection. arXiv preprint arXiv:2211.13529 2022.
33	R2Pfusion-Det			95.75 %	96.53 %	93.12 %	0.3 s	1 core @ 2.5 Ghz (C/C++)

34	GLENet-VR		code	95.73 %	96.84 %	90.80 %	0.04 s	1 core @ 2.5 Ghz (C/C++)
Y. Zhang, Q. Zhang, Z. Zhu, J. Hou and Y. Yuan: GLENet: Boosting 3D object detectors with generative label uncertainty estimation. International Journal of Computer Vision 2023. Y. Zhang, J. Hou and Y. Yuan: A Comprehensive Study of the Robustness for LiDAR-based 3D Object Detectors against Adversarial Attacks. International Journal of Computer Vision 2023.
35	GraR-Pi		code	95.72 %	98.57 %	92.55 %	0.03 s	1 core @ 2.5 Ghz (C/C++)
H. Yang, Z. Liu, X. Wu, W. Wang, W. Qian, X. He and D. Cai: Graph R-CNN: Towards Accurate 3D Object Detection with Semantic-Decorated Local Graph. ECCV 2022.
36	OcTr			95.69 %	96.44 %	90.78 %	0.06 s	GPU @ 2.5 Ghz (Python + C/C++)
C. Zhou, Y. Zhang, J. Chen and D. Huang: OcTr: Octree-based Transformer for 3D Object Detection. CVPR 2023.
37	mm3d			95.65 %	96.84 %	90.69 %	0.2 s	1 core @ 2.5 Ghz (C/C++)

38	DVF-V			95.63 %	96.59 %	90.71 %	0.1 s	1 core @ 2.5 Ghz (C/C++)
A. Mahmoud, J. Hu and S. Waslander: Dense Voxel Fusion for 3D Object Detection. WACV 2023.
39	DSGN++		code	95.58 %	98.04 %	88.09 %	0.2 s	GeForce RTX 2080Ti
Y. Chen, S. Huang, S. Liu, B. Yu and J. Jia: DSGN++: Exploiting Visual-Spatial Relation for Stereo-Based 3D Detectors. IEEE Transactions on Pattern Analysis and Machine Intelligence 2022.
40	Fast-CLOCs			95.57 %	96.66 %	90.70 %	0.1 s	GPU @ 2.5 Ghz (Python)
S. Pang, D. Morris and H. Radha: Fast-CLOCs: Fast Camera-LiDAR Object Candidates Fusion for 3D Object Detection. Proceedings of the IEEE/CVF Winter Conference on Applications of Computer Vision (WACV) 2022.
41	TSSTDet			95.56 %	96.54 %	92.71 %	0.08 s	1 core @ 2.5 Ghz (C/C++)
H. Hoang, D. Bui and M. Yoo: TSSTDet: Transformation-Based 3-D Object Detection via a Spatial Shape Transformer. IEEE Sensors Journal 2024.
42	LCANet			95.55 %	96.55 %	92.88 %	1 s	1 core @ 2.5 Ghz (C/C++)

43	3D HANet		code	95.54 %	98.59 %	92.66 %	0.1 s	1 core @ 2.5 Ghz (C/C++)
Q. Xia, Y. Chen, G. Cai, G. Chen, D. Xie, J. Su and Z. Wang: 3D HANet: A Flexible 3D Heatmap Auxiliary Network for Object Detection. IEEE Transactions on Geoscience and Remote Sensing 2023.
44	FARP-Net		code	95.53 %	96.10 %	92.98 %	0.06 s	GPU @ 2.5 Ghz (Python + C/C++)
T. Xie, L. Wang, K. Wang, R. Li, X. Zhang, H. Zhang, L. Yang, H. Liu and J. Li: FARP-Net: Local-Global Feature Aggregation and Relation-Aware Proposals for 3D Object Detection. IEEE Transactions on Multimedia 2023.
45	CasA		code	95.53 %	96.51 %	92.71 %	0.1 s	1 core @ 2.5 Ghz (C/C++)
H. Wu, J. Deng, C. Wen, X. Li and C. Wang: CasA: A Cascade Attention Network for 3D Object Detection from LiDAR point clouds. IEEE Transactions on Geoscience and Remote Sensing 2022.
46	mat3D			95.48 %	98.79 %	92.73 %	0.2 s	1 core @ 2.5 Ghz (C/C++)

47	LoGoNet		code	95.44 %	96.59 %	92.89 %	0.1 s	1 core @ 2.5 Ghz (C/C++)
X. Li, T. Ma, Y. Hou, B. Shi, Y. Yang, Y. Liu, X. Wu, Q. Chen, Y. Li, Y. Qiao and others: LoGoNet: Towards Accurate 3D Object Detection with Local-to-Global Cross-Modal Fusion. CVPR 2023.
48	TRTConv-L			95.41 %	96.37 %	92.57 %	0.01 s	1 core @ 2.5 Ghz (C/C++)
ERROR: Wrong syntax in BIBTEX file.
49	ECA			95.40 %	98.10 %	92.62 %	0.08 s	GPU @ 1.5 Ghz (Python)

50	GD-MAE			95.36 %	98.31 %	90.19 %	0.07 s	1 core @ 2.5 Ghz (Python + C/C++)
H. Yang, T. He, J. Liu, H. Chen, B. Wu, B. Lin, X. He and W. Ouyang: GD-MAE: Generative Decoder for MAE Pre- training on LiDAR Point Clouds. CVPR 2023.
51	HS-fusion			95.36 %	98.30 %	92.45 %	- s	1 core @ 2.5 Ghz (Python)
ERROR: Wrong syntax in BIBTEX file.
52	DVF-PV			95.35 %	96.40 %	92.37 %	0.1 s	1 core @ 2.5 Ghz (C/C++)
A. Mahmoud, J. Hu and S. Waslander: Dense Voxel Fusion for 3D Object Detection. WACV 2023.
53	BADet		code	95.34 %	98.65 %	90.28 %	0.14 s	1 core @ 2.5 Ghz (C/C++)
R. Qian, X. Lai and X. Li: BADet: Boundary-Aware 3D Object Detection from Point Clouds. Pattern Recognition 2022.
54	BVPConv-T			95.30 %	96.26 %	92.65 %	0.05 s	1 core @ 2.5 Ghz (Python + C/C++)
ERROR: Wrong syntax in BIBTEX file.
55	SASA		code	95.29 %	96.00 %	92.42 %	0.04 s	1 core @ 2.5 Ghz (Python + C/C++)
C. Chen, Z. Chen, J. Zhang and D. Tao: SASA: Semantics-Augmented Set Abstraction for Point-based 3D Object Detection. arXiv preprint arXiv:2201.01976 2022.
56	PG-RCNN		code	95.27 %	96.64 %	90.37 %	0.06 s	GPU @ 1.5 Ghz (Python)
I. Koo, I. Lee, S. Kim, H. Kim, W. Jeon and C. Kim: PG-RCNN: Semantic Surface Point Generation for 3D Object Detection. 2023.
57	c2f			95.27 %	96.23 %	92.45 %	1 s	1 core @ 2.5 Ghz (C/C++)

58	FIRM-Net_SCF+			95.23 %	96.29 %	92.44 %	0.07 s	1 core @ 2.5 Ghz (Python)

59	Focals Conv		code	95.23 %	96.29 %	92.60 %	0.1 s	1 core @ 2.5 Ghz (C/C++)
Y. Chen, Y. Li, X. Zhang, J. Sun and J. Jia: Focal Sparse Convolutional Networks for 3D Object Detection. Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition 2022.
60	FIRM-Net-SCF			95.21 %	96.28 %	92.42 %	0.07 s	1 core @ 2.5 Ghz (Python)

61	EQ-PVRCNN		code	95.20 %	98.22 %	92.47 %	0.2 s	GPU @ 2.5 Ghz (Python + C/C++)
Z. Yang, L. Jiang, Y. Sun, B. Schiele and J. Jia: A Unified Query-based Paradigm for Point Cloud Understanding. Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition 2022.
62	SCDA-Net			95.19 %	98.57 %	92.64 %	- s	1 core @ 2.5 Ghz (C/C++)

63	CasA++		code	95.17 %	95.81 %	94.10 %	0.1 s	1 core @ 2.5 Ghz (C/C++)
H. Wu, J. Deng, C. Wen, X. Li and C. Wang: CasA: A Cascade Attention Network for 3D Object Detection from LiDAR point clouds. IEEE Transactions on Geoscience and Remote Sensing 2022.
64	SE-SSD		code	95.17 %	96.55 %	90.00 %	0.03 s	1 core @ 2.5 Ghz (Python + C/C++)
W. Zheng, W. Tang, L. Jiang and C. Fu: SE-SSD: Self-Ensembling Single-Stage Object Detector From Point Cloud. CVPR 2021.
65	TRTConv-T			95.15 %	98.35 %	92.45 %	0.04 s	1 core @ 2.5 Ghz (Python + C/C++)
ERROR: Wrong syntax in BIBTEX file.
66	VoxSeT		code	95.13 %	96.15 %	90.38 %	33 ms	1 core @ 2.5 Ghz (C/C++)
C. He, R. Li, S. Li and L. Zhang: Voxel Set Transformer: A Set-to-Set Approach to 3D Object Detection from Point Clouds. CVPR 2022.
67	HMFI		code	95.05 %	96.28 %	92.28 %	0.1 s	1 core @ 2.5 Ghz (C/C++)
X. Li, B. Shi, Y. Hou, X. Wu, T. Ma, Y. Li and L. He: Homogeneous Multi-modal Feature Fusion and Interaction for 3D Object Detection. ECCV 2022.
68	RagNet3D		code	95.04 %	96.26 %	92.48 %	0.05 s	1 core @ 2.5 Ghz (C/C++)
J. Chen, Y. Han, Z. Yan, J. Qian, J. Li and J. Yang: Ragnet3d: Learning Distinguishable Representation for Pooled Grids in 3d Object Detection. Available at SSRN 4979473 .
69	voxel-rcnn+++		code	95.03 %	96.37 %	92.07 %	0.08 s	GPU @ 2.5 Ghz (Python)

70	SPANet			95.03 %	96.31 %	89.99 %	0.06 s	1 core @ 2.5 Ghz (C/C++)
Y. Ye: SPANet: Spatial and Part-Aware Aggregation Network for 3D Object Detection. Pacific Rim International Conference on Artificial Intelligence 2021.
71	Pyramid R-CNN			95.03 %	95.87 %	92.46 %	0.07 s	1 core @ 2.5 Ghz (C/C++)
J. Mao, M. Niu, H. Bai, X. Liang, H. Xu and C. Xu: Pyramid R-CNN: Towards Better Performance and Adaptability for 3D Object Detection. ICCV 2021.
72	MPC3DNet			95.03 %	98.48 %	92.24 %	0.05 s	GPU @ 1.5 Ghz (Python)

73	VPFNet		code	95.01 %	96.03 %	92.41 %	0.2 s	1 core @ 2.5 Ghz (C/C++)
C. Wang, H. Chen and L. Fu: VPFNet: Voxel-Pixel Fusion Network for Multi-class 3D Object Detection. 2021. C. Wang, H. Chen, Y. Chen, P. Hsiao and L. Fu: VoPiFNet: Voxel-Pixel Fusion Network for Multi-Class 3D Object Detection. IEEE Transactions on Intelligent Transportation Systems 2024.
74	EPNet++			95.00 %	96.70 %	91.82 %	0.1 s	GPU @ 2.5 Ghz (Python)
Z. Liu, T. Huang, B. Li, X. Chen, X. Wang and X. Bai: EPNet++: Cascade Bi-Directional Fusion for Multi-Modal 3D Object Detection. IEEE Transactions on Pattern Analysis and Machine Intelligence 2022.
75	CAIA_PRO		code	95.00 %	95.71 %	90.31 %	0.01 s	1 core @ 2.5 Ghz (C/C++)

76	USVLab BSAODet		code	94.99 %	96.23 %	92.36 %	0.04 s	1 core @ 2.5 Ghz (C/C++)
W. Xiao, Y. Peng, C. Liu, J. Gao, Y. Wu and X. Li: Balanced Sample Assignment and Objective for Single-Model Multi-Class 3D Object Detection. IEEE Transactions on Circuits and Systems for Video Technology 2023.
77	Voxel R-CNN		code	94.96 %	96.47 %	92.24 %	0.04 s	GPU @ 3.0 Ghz (C/C++)
J. Deng, S. Shi, P. Li, W. Zhou, Y. Zhang and H. Li: Voxel R-CNN: Towards High Performance Voxel-based 3D Object Detection . AAAI 2021.
78	BVPConv-L			94.92 %	96.28 %	92.26 %	0.01 s	1 core @ 2.5 Ghz (Python + C/C++)
ERROR: Wrong syntax in BIBTEX file.
79	PDV		code	94.91 %	96.06 %	92.30 %	0.1 s	1 core @ 2.5 Ghz (C/C++)
J. Hu, T. Kuai and S. Waslander: Point Density-Aware Voxels for LiDAR 3D Object Detection. CVPR 2022.
80	SIENet		code	94.85 %	96.01 %	92.23 %	0.08 s	1 core @ 2.5 Ghz (Python)
Z. Li, Y. Yao, Z. Quan, W. Yang and J. Xie: SIENet: Spatial Information Enhancement Network for 3D Object Detection from Point Cloud. 2021.
81	SQD		code	94.85 %	98.20 %	92.26 %	0.06 s	1 core @ 2.5 Ghz (C/C++)
Y. Mo, Y. Wu, J. Zhao, Z. Hou, W. Huang, Y. Hu, J. Wang and J. Yan: Sparse Query Dense: Enhancing 3D Object Detection with Pseudo Points. ACM MM Oral 2024.
82	SVFMamba		code	94.84 %	95.51 %	92.03 %	N/A s	1 core @ 2.5 Ghz (C/C++)

83	VoTr-TSD		code	94.81 %	95.95 %	92.24 %	0.07 s	1 core @ 2.5 Ghz (C/C++)
J. Mao, Y. Xue, M. Niu, H. Bai, J. Feng, X. Liang, H. Xu and C. Xu: Voxel Transformer for 3D Object Detection. ICCV 2021.
84	GraphAlign(ICCV2023)		code	94.79 %	98.04 %	92.35 %	0.03 s	GPU @ 2.0 Ghz (Python)
Z. Song, H. Wei, L. Bai, L. Yang and C. Jia: GraphAlign: Enhancing accurate feature alignment by graph matching for multi-modal 3D object detection. Proceedings of the IEEE/CVF International Conference on Computer Vision 2023.
85	L-AUG			94.76 %	95.80 %	91.94 %	0.1 s	1 core @ 2.5 Ghz (C/C++)
T. Cortinhal, I. Gouigah and E. Aksoy: Semantics-aware LiDAR-Only Pseudo Point Cloud Generation for 3D Object Detection. 2023.
86	focal			94.75 %	95.88 %	92.36 %	100 s	1 core @ 2.5 Ghz (Python)

87	GEFPN			94.75 %	95.88 %	92.36 %	0.5 s	1 core @ 2.5 Ghz (C/C++)

88	GeVo			94.75 %	95.88 %	92.36 %	0.05 s	1 core @ 2.5 Ghz (C/C++)

89	M3DeTR		code	94.70 %	97.37 %	91.89 %	n/a s	GPU @ 1.0 Ghz (Python)
T. Guan, J. Wang, S. Lan, R. Chandra, Z. Wu, L. Davis and D. Manocha: M3DeTR: Multi-representation, Multi- scale, Mutual-relation 3D Object Detection with Transformers. 2021.
90	MonoSample (DID-M3D)		code	94.69 %	96.30 %	85.10 %	0.2 s	1 core @ 2.5 Ghz (C/C++)
J. Qiao, B. Liu, J. Yang, B. Wang, S. Xiu, X. Du and X. Nie: MonoSample: Synthetic 3D Data Augmentation Method in Monocular 3D Object Detection. IEEE Robotics and Automation Letters 2024.
91	MLFusion-VS			94.68 %	98.25 %	92.25 %	0.06 s	1 core @ 2.5 Ghz (C/C++)

92	XView			94.66 %	95.88 %	92.07 %	0.1 s	1 core @ 2.5 Ghz (C/C++)
L. Xie, G. Xu, D. Cai and X. He: X-view: Non-egocentric Multi-View 3D Object Detector. 2021.
93	SP_SECOND_IOU		code	94.64 %	95.84 %	92.07 %	0.04 s	1 core @ 2.5 Ghz (Python)

94	StructuralIF			94.64 %	96.12 %	91.85 %	0.02 s	8 cores @ 2.5 Ghz (Python)
J. Pei An: Deep structural information fusion for 3D object detection on LiDAR-camera system. Accepted in CVIU 2021.
95	AFFN			94.63 %	95.89 %	92.00 %	0.5 s	1 core @ 2.5 Ghz (C/C++)

96	AFFN-G			94.62 %	98.05 %	92.13 %	0.5 s	1 core @ 2.5 Ghz (C/C++)

97	test			94.61 %	97.49 %	91.75 %	0.04 s	GPU @ 1.5 Ghz (Python + C/C++)

98	P2V-RCNN			94.59 %	96.01 %	92.13 %	0.1 s	1 core @ 2.5 Ghz (C/C++)
J. Li, S. Luo, Z. Zhu, H. Dai, A. Krylov, Y. Ding and L. Shao: P2V-RCNN: Point to Voxel Feature Learning for 3D Object Detection from Point Clouds. IEEE Access 2021.
99	CAT-Det			94.57 %	95.95 %	91.88 %	0.3 s	GPU @ 2.5 Ghz (Python + C/C++)
Y. Zhang, J. Chen and D. Huang: CAT-Det: Contrastively Augmented Transformer for Multi-modal 3D Object Detection. CVPR 2022.
100	MMLab PV-RCNN		code	94.57 %	98.15 %	91.85 %	0.08 s	1 core @ 2.5 Ghz (Python + C/C++)
S. Shi, C. Guo, L. Jiang, Z. Wang, J. Shi, X. Wang and H. Li: PV-RCNN: Point-Voxel Feature Set Abstraction for 3D Object Detection. CVPR 2020.
101	CG-SSD			94.53 %	95.95 %	91.92 %	0.01 s	1 core @ 2.5 Ghz (C/C++)

102	DSA-PV-RCNN		code	94.52 %	95.84 %	91.93 %	0.08 s	1 core @ 2.5 Ghz (Python + C/C++)
P. Bhattacharyya, C. Huang and K. Czarnecki: SA-Det3D: Self-Attention Based Context-Aware 3D Object Detection. 2021.
103	RangeDet (Official)		code	94.51 %	95.48 %	91.57 %	0.02 s	1 core @ 2.5 Ghz (C/C++)
L. Fan, X. Xiong, F. Wang, N. Wang and Z. Zhang: RangeDet: In Defense of Range View for LiDAR-Based 3D Object Detection. Proceedings of the IEEE/CVF International Conference on Computer Vision (ICCV) 2021.
104	second_iou_baseline		code	94.51 %	95.73 %	91.80 %	0.05 s	1 core @ 2.5 Ghz (Python)

105	second_iou_baseline			94.48 %	95.74 %	91.80 %	0.03 s	1 core @ 2.5 Ghz (Python)

106	MVRA + I-FRCNN+			94.46 %	95.66 %	81.74 %	0.18 s	GPU @ 2.5 Ghz (Python)
H. Choi, H. Kang and Y. Hyun: Multi-View Reprojection Architecture for Orientation Estimation. The IEEE International Conference on Computer Vision (ICCV) Workshops 2019.
107	SVGA-Net			94.45 %	96.02 %	91.54 %	0.03s	1 core @ 2.5 Ghz (Python + C/C++)
Q. He, Z. Wang, H. Zeng, Y. Zeng and Y. Liu: SVGA-Net: Sparse Voxel-Graph Attention Network for 3D Object Detection from Point Clouds. AAAI 2022.
108	PASS-PV-RCNN-Plus			94.45 %	95.77 %	91.89 %	1 s	1 core @ 2.5 Ghz (Python)
Anonymous: Leveraging Anchor-based LiDAR 3D Object Detection via Point Assisted Sample Selection. will submit to computer vision conference/journal 2024.
109	DVFENet			94.44 %	95.33 %	91.55 %	0.05 s	1 core @ 2.5 Ghz (Python + C/C++)
Y. He, G. Xia, Y. Luo, L. Su, Z. Zhang, W. Li and P. Wang: DVFENet: Dual-branch Voxel Feature Extraction Network for 3D Object Detection. Neurocomputing 2021.
110	RangeIoUDet			94.42 %	95.69 %	91.70 %	0.02 s	GPU @ 2.5 Ghz (Python + C/C++)
Z. Liang, Z. Zhang, M. Zhang, X. Zhao and S. Pu: RangeIoUDet: Range Image Based Real-Time 3D Object Detector Optimized by Intersection Over Union. CVPR 2021.
111	HA-PillarNet			94.40 %	98.02 %	92.00 %	0.05 s	1 core @ 2.5 Ghz (Python)

112	BVIFusion+			94.40 %	95.77 %	91.61 %	0.09 s	1 core @ 2.5 Ghz (Python)

113	SFA_IGCL_Focalsconv*		code	94.35 %	95.90 %	92.02 %	0.2 s	1 core @ 2.5 Ghz (C/C++)

114	MFB3D			94.34 %	97.62 %	91.59 %	0.14 s	1 core @ 2.5 Ghz (Python)

115	AFFN-Ga			94.31 %	95.78 %	91.85 %	0.5 s	1 core @ 2.5 Ghz (C/C++)

116	SERCNN			94.24 %	96.31 %	89.71 %	0.1 s	1 core @ 2.5 Ghz (Python)
D. Zhou, J. Fang, X. Song, L. Liu, J. Yin, Y. Dai, H. Li and R. Yang: Joint 3D Instance Segmentation and Object Detection for Autonomous Driving. Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition 2020.
117	EPNet		code	94.22 %	96.13 %	89.68 %	0.1 s	1 core @ 2.5 Ghz (Python + C/C++)
T. Huang, Z. Liu, X. Chen and X. Bai: EPNet: Enhancing Point Features with Image Semantics for 3D Object Detection. ECCV 2020.
118	dsvd+vx			94.20 %	95.08 %	91.35 %	0.1 s	1 core @ 2.5 Ghz (C/C++)

119	SRDL			94.08 %	95.83 %	91.55 %	0.05 s	1 core @ 2.5 Ghz (Python + C/C++)
ERROR: Wrong syntax in BIBTEX file.
120	CGML			94.08 %	97.55 %	91.74 %	0.33 s	1 core @ 2.5 Ghz (C/C++)

121	Voxel RCNN-Focal*		code	94.06 %	95.61 %	91.83 %	0.2 s	1 core @ 2.5 Ghz (Python)

122	FocalsConv*			94.01 %	97.66 %	91.72 %	0.2 s	1 core @ 2.5 Ghz (C/C++)

123	HMNet			93.96 %	95.50 %	91.05 %	0.1 s	1 core @ 2.5 Ghz (C/C++)

124	Res3DNet			93.95 %	95.43 %	91.15 %	0.05 s	GPU @ 3.5 Ghz (Python)

125	RangeRCNN			93.90 %	95.47 %	91.53 %	0.06 s	GPU @ 2.5 Ghz (Python + C/C++)
Z. Liang, M. Zhang, Z. Zhang, X. Zhao and S. Pu: RangeRCNN: Towards Fast and Accurate 3D Object Detection with Range Image Representation. arXiv preprint arXiv:2009.00206 2020.
126	SA V1			93.88 %	94.84 %	90.92 %	0.5 s	GPU @ 2.5 Ghz (Python)

127	MSMA V1			93.84 %	95.19 %	90.71 %	0.5 s	GPU @ 2.5 Ghz (Python)

128	SIF			93.79 %	95.48 %	91.30 %	0.1 s	1 core @ 2.5 Ghz (C/C++)
P. An: SIF. Submitted to CVIU 2021.
129	SFA-GCL(80)		code	93.78 %	96.89 %	90.74 %	0.04 s	1 core @ 2.5 Ghz (C/C++)

130	DD3D		code	93.78 %	94.67 %	88.99 %	n/a s	1 core @ 2.5 Ghz (C/C++)
D. Park, R. Ambrus, V. Guizilini, J. Li and A. Gaidon: Is Pseudo-Lidar needed for Monocular 3D Object detection?. IEEE/CVF International Conference on Computer Vision (ICCV) .
131	MGAF-3DSSD		code	93.77 %	94.45 %	86.25 %	0.1 s	1 core @ 2.5 Ghz (Python)
J. Li, H. Dai, L. Shao and Y. Ding: Anchor-free 3D Single Stage Detector with Mask-Guided Attention for Point Cloud. MM '21: The 29th ACM International Conference on Multimedia (ACM MM) 2021.
132	SFA-GCL		code	93.75 %	96.86 %	90.71 %	0.04 s	1 core @ 2.5 Ghz (C/C++)

133	SFA-GCL(80, k=4)		code	93.72 %	96.84 %	90.67 %	0.04 s	1 core @ 2.5 Ghz (C/C++)

134	MMLAB LIGA-Stereo		code	93.71 %	96.40 %	86.00 %	0.4 s	1 core @ 2.5 Ghz (Python + C/C++)
X. Guo, S. Shi, X. Wang and H. Li: LIGA-Stereo: Learning LiDAR Geometry Aware Representations for Stereo-based 3D Detector. Proceedings of the IEEE/CVF International Conference on Computer Vision (ICCV) 2021.
135	Sem-Aug			93.69 %	96.78 %	88.69 %	0.1 s	GPU @ 2.5 Ghz (Python)
L. Zhao, M. Wang and Y. Yue: Sem-Aug: Improving Camera-LiDAR Feature Fusion With Semantic Augmentation for 3D Vehicle Detection. IEEE Robotics and Automation Letters 2022.
136	3ONet			93.58 %	96.86 %	88.45 %	0.1 s	1 core @ 2.5 Ghz (C/C++)
H. Hoang and M. Yoo: 3ONet: 3-D Detector for Occluded Object Under Obstructed Conditions. IEEE Sensors Journal 2023.
137	Patches - EMP			93.58 %	97.88 %	90.31 %	0.5 s	GPU @ 2.5 Ghz (Python)
J. Lehner, A. Mitterecker, T. Adler, M. Hofmarcher, B. Nessler and S. Hochreiter: Patch Refinement: Localized 3D Object Detection. arXiv preprint arXiv:1910.04093 2019.
138	SFA-GCL(baseline)		code	93.58 %	96.81 %	90.54 %	0.04 s	1 core @ 2.5 Ghz (C/C++)

139	SFA-GCL_dataaug		code	93.57 %	96.72 %	90.49 %	0.04 s	1 core @ 2.5 Ghz (C/C++)

140	SFA-GCL		code	93.57 %	96.84 %	90.47 %	0.04 s	1 core @ 2.5 Ghz (C/C++)

141	PA3DNet			93.55 %	96.56 %	88.56 %	0.1 s	GPU @ 2.5 Ghz (Python + C/C++)
M. Wang, L. Zhao and Y. Yue: PA3DNet: 3-D Vehicle Detection with Pseudo Shape Segmentation and Adaptive Camera- LiDAR Fusion. IEEE Transactions on Industrial Informatics 2023.
142	MVAF-Net		code	93.54 %	95.35 %	90.70 %	0.06 s	1 core @ 2.5 Ghz (Python + C/C++)
G. Wang, B. Tian, Y. Zhang, L. Chen, D. Cao and J. Wu: Multi-View Adaptive Fusion Network for 3D Object Detection. arXiv preprint arXiv:2011.00652 2020.
143	IA-SSD (multi)		code	93.47 %	96.07 %	90.51 %	0.014 s	1 core @ 2.5 Ghz (C/C++)
Y. Zhang, Q. Hu, G. Xu, Y. Ma, J. Wan and Y. Guo: Not All Points Are Equal: Learning Highly Efficient Point-based Detectors for 3D LiDAR Point Clouds. CVPR 2022.
144	SCNet3D			93.44 %	96.53 %	90.71 %	0.08 s	1 core @ 2.5 Ghz (Python)

145	IA-SSD (single)		code	93.41 %	96.23 %	88.34 %	0.013 s	1 core @ 2.5 Ghz (C/C++)
Y. Zhang, Q. Hu, G. Xu, Y. Ma, J. Wan and Y. Guo: Not All Points Are Equal: Learning Highly Efficient Point-based Detectors for 3D LiDAR Point Clouds. CVPR 2022.
146	CIA-SSD		code	93.34 %	96.65 %	85.76 %	0.03 s	1 core @ 2.5 Ghz (Python + C/C++)
W. Zheng, W. Tang, S. Chen, L. Jiang and C. Fu: CIA-SSD: Confident IoU-Aware Single-Stage Object Detector From Point Cloud. AAAI 2021.
147	SeSame-point		code	93.32 %	95.20 %	90.14 %	N/A s	TITAN RTX @ 1.35 Ghz (Python)
H. O, C. Yang and K. Huh: SeSame: Simple, Easy 3D Object Detection with Point-Wise Semantics. Proceedings of the Asian Conference on Computer Vision (ACCV) 2024.
148	Deep MANTA			93.31 %	98.83 %	82.95 %	0.7 s	GPU @ 2.5 Ghz (Python + C/C++)
F. Chabot, M. Chaouch, J. Rabarisoa, C. Teulière and T. Chateau: Deep MANTA: A Coarse-to-fine Many-Task Network for joint 2D and 3D vehicle analysis from monocular image. CVPR 2017.
149	StereoDistill			93.29 %	97.57 %	87.48 %	0.4 s	1 core @ 2.5 Ghz (Python)
Z. Liu, X. Ye, X. Tan, D. Errui, Y. Zhou and X. Bai: StereoDistill: Pick the Cream from LiDAR for Distilling Stereo-based 3D Object Detection. Proceedings of the AAAI Conference on Artificial Intelligence 2023.
150	GSG-FPS		code	93.28 %	95.72 %	90.33 %	0.01 s	1 core @ 2.5 Ghz (C/C++)

151	LPCG-Monoflex		code	93.26 %	96.68 %	83.34 %	0.03 s	1 core @ 2.5 Ghz (C/C++)
L. Peng, F. Liu, Z. Yu, S. Yan, D. Deng, Z. Yang, H. Liu and D. Cai: Lidar Point Cloud Guided Monocular 3D Object Detection. ECCV 2022.
152	MonoLiG		code	93.23 %	96.56 %	83.42 %	0.03 s	1 core @ 2.5 Ghz (C/C++)
A. Hekimoglu, M. Schmidt and A. Ramiro: Monocular 3D Object Detection with LiDAR Guided Semi Supervised Active Learning. 2023.
153	CityBrainLab-CT3D		code	93.20 %	96.26 %	90.44 %	0.07 s	1 core @ 2.5 Ghz (Python + C/C++)
H. Sheng, S. Cai, Y. Liu, B. Deng, J. Huang, X. Hua and M. Zhao: Improving 3D Object Detection with Channel- wise Transformer. ICCV 2021.
154	DFAF3D			93.20 %	96.54 %	90.03 %	0.05 s	1 core @ 2.5 Ghz (Python)
Q. Tang, X. Bai, J. Guo, B. Pan and W. Jiang: DFAF3D: A dual-feature-aware anchor-free single-stage 3D detector for point clouds. Image and Vision Computing 2023.
155	Test_dif		code	93.16 %	95.77 %	90.26 %	0.01 s	1 core @ 2.5 Ghz (C/C++)

156	MonoLSS			93.11 %	95.99 %	83.14 %	0.04 s	1 core @ 2.5 Ghz (Python)
Z. Li, J. Jia and Y. Shi: MonoLSS: Learnable Sample Selection For Monocular 3D Detection. International Conference on 3D Vision 2024.
157	MonoAFKD			93.11 %	95.99 %	83.13 %	0.03 s	1 core @ 2.5 Ghz (C/C++)

158	SNVC		code	93.09 %	96.27 %	85.51 %	1 s	GPU @ 1.0 Ghz (Python)
S. Li, Z. Liu, Z. Shen and K. Cheng: Stereo Neural Vernier Caliper. Proceedings of the AAAI Conference on Artificial Intelligence 2022.
159	H^23D R-CNN		code	93.03 %	96.13 %	90.33 %	0.03 s	1 core @ 2.5 Ghz (C/C++)
J. Deng, W. Zhou, Y. Zhang and H. Li: From Multi-View to Hollow-3D: Hallucinated Hollow-3D R-CNN for 3D Object Detection. IEEE Transactions on Circuits and Systems for Video Technology 2021.
160	AM			93.01 %	96.13 %	85.21 %	0.04 s	1 core @ 2.5 Ghz (C/C++)

161	FromVoxelToPoint		code	92.98 %	96.07 %	90.40 %	0.1 s	1 core @ 2.5 Ghz (C/C++)
J. Li, H. Dai, L. Shao and Y. Ding: From Voxel to Point: IoU-guided 3D Object Detection for Point Cloud with Voxel-to- Point Decoder. MM '21: The 29th ACM International Conference on Multimedia (ACM MM) 2021.
162	EBM3DOD		code	92.88 %	96.39 %	87.58 %	0.12 s	1 core @ 2.5 Ghz (Python)
F. Gustafsson, M. Danelljan and T. Schön: Accurate 3D Object Detection using Energy- Based Models. arXiv preprint arXiv:2012.04634 2020.
163	Struc info fusion II			92.88 %	96.44 %	87.67 %	0.05 s	GPU @ 2.5 Ghz (Python)
P. An, J. Liang, J. Ma, K. Yu and B. Fang: Struc info fusion. Submitted to CVIU 2021.
164	HotSpotNet			92.74 %	96.20 %	89.68 %	0.04 s	1 core @ 2.5 Ghz (Python + C/C++)
Q. Chen, L. Sun, Z. Wang, K. Jia and A. Yuille: object as hotspots. Proceedings of the European Conference on Computer Vision (ECCV) 2020.
165	Struc info fusion I			92.71 %	96.24 %	87.55 %	0.05 s	1 core @ 2.5 Ghz (Python)
P. An, J. Liang, J. Ma, K. Yu and B. Fang: Struc info fusion. Submitted to CVIU 2021.
166	EBM3DOD baseline		code	92.70 %	96.31 %	87.44 %	0.05 s	1 core @ 2.5 Ghz (Python)
F. Gustafsson, M. Danelljan and T. Schön: Accurate 3D Object Detection using Energy- Based Models. arXiv preprint arXiv:2012.04634 2020.
167	MonoCD		code	92.65 %	96.36 %	85.17 %	n/a s	1 core @ 2.5 Ghz (Python)
L. Yan, P. Yan, S. Xiong, X. Xiang and Y. Tan: MonoCD: Monocular 3D Object Detection with Complementary Depths. CVPR 2024.
168	Fade-kd			92.59 %	96.18 %	89.54 %	0.01 s	1 core @ 2.5 Ghz (C/C++)

169	SARPNET			92.58 %	95.82 %	87.33 %	0.05 s	1 core @ 2.5 Ghz (Python + C/C++)
Y. Ye, H. Chen, C. Zhang, X. Hao and Z. Zhang: SARPNET: Shape Attention Regional Proposal Network for LiDAR-based 3D Object Detection. Neurocomputing 2019.
170	Patches			92.57 %	96.31 %	87.41 %	0.15 s	GPU @ 2.0 Ghz
J. Lehner, A. Mitterecker, T. Adler, M. Hofmarcher, B. Nessler and S. Hochreiter: Patch Refinement: Localized 3D Object Detection. arXiv preprint arXiv:1910.04093 2019.
171	R-GCN			92.53 %	96.16 %	87.45 %	0.16 s	GPU @ 2.5 Ghz (Python)
J. Zarzar, S. Giancola and B. Ghanem: PointRGCN: Graph Convolution Networks for 3D Vehicles Detection Refinement. ArXiv 2019.
172	PI-RCNN			92.52 %	96.15 %	87.47 %	0.1 s	1 core @ 2.5 Ghz (Python)
L. Xie, C. Xiang, Z. Yu, G. Xu, Z. Yang, D. Cai and X. He: PI-RCNN: An Efficient Multi-sensor 3D Object Detector with Point-based Attentive Cont-conv Fusion Module. AAAI 2020 : The Thirty-Fourth AAAI Conference on Artificial Intelligence 2020.
173	CenterNet3D			92.48 %	95.71 %	89.54 %	0.04 s	GPU @ 1.5 Ghz (Python)
G. Wang, B. Tian, Y. Ai, T. Xu, L. Chen and D. Cao: CenterNet3D:An Anchor free Object Detector for Autonomous Driving. 2020.
174	zqd_test			92.46 %	95.85 %	86.73 %	0.2 s	1 core @ 2.5 Ghz (Python)

175	PointPainting			92.43 %	98.36 %	89.49 %	0.4 s	GPU @ 2.5 Ghz (Python + C/C++)
S. Vora, A. Lang, B. Helou and O. Beijbom: PointPainting: Sequential Fusion for 3D Object Detection. CVPR 2020.
176	3D IoU-Net			92.42 %	96.31 %	87.60 %	0.1 s	1 core @ 2.5 Ghz (C/C++)
J. Li, S. Luo, Z. Zhu, H. Dai, S. Krylov, Y. Ding and L. Shao: 3D IoU-Net: IoU Guided 3D Object Detector for Point Clouds. arXiv preprint arXiv:2004.04962 2020.
177	R50_SACINet			92.41 %	95.43 %	89.40 %	0.06 s	1 core @ 2.5 Ghz (C/C++)

178	CLOCs_SecCas			92.37 %	95.16 %	88.43 %	0.1 s	1 core @ 2.5 Ghz (Python)
S. Pang, D. Morris and H. Radha: CLOCs: Camera-LiDAR Object Candidates Fusion for 3D Object Detection. 2020 IEEE/RSJ International Conference on Intelligent Robots and Systems (IROS) 2020.
179	ACDet		code	92.36 %	96.07 %	89.18 %	0.05 s	1 core @ 2.5 Ghz (C/C++)
J. Xu, G. Wang, X. Zhang and G. Wan: ACDet: Attentive Cross-view Fusion for LiDAR-based 3D Object Detection. 3DV 2022.
180	L_SACINet			92.30 %	95.39 %	87.34 %	0.04 s	1 core @ 2.5 Ghz (C/C++)

181	DASS			92.25 %	96.20 %	87.26 %	0.09 s	1 core @ 2.0 Ghz (Python)
O. Unal, L. Van Gool and D. Dai: Improving Point Cloud Semantic Segmentation by Learning 3D Object Detection. Proceedings of the IEEE/CVF Winter Conference on Applications of Computer Vision (WACV) 2021.
182	S-AT GCN			92.24 %	95.02 %	90.46 %	0.02 s	GPU @ 2.0 Ghz (Python)
L. Wang, C. Wang, X. Zhang, T. Lan and J. Li: S-AT GCN: Spatial-Attention Graph Convolution Network based Feature Enhancement for 3D Object Detection. CoRR 2021.
183	Fade			92.23 %	97.62 %	87.09 %	0.01 s	1 core @ 2.5 Ghz (C/C++)

184	Sem-Aug-PointRCNN++			92.20 %	95.64 %	87.48 %	0.1 s	8 cores @ 3.0 Ghz (Python)
L. Zhao, M. Wang and Y. Yue: Sem-Aug: Improving Camera-LiDAR Feature Fusion With Semantic Augmentation for 3D Vehicle Detection. IEEE Robotics and Automation Letters 2022.
185	SegVoxelNet			92.16 %	95.86 %	86.90 %	0.04 s	1 core @ 2.5 Ghz (Python)
H. Yi, S. Shi, M. Ding, J. Sun, K. Xu, H. Zhou, Z. Wang, S. Li and G. Wang: SegVoxelNet: Exploring Semantic Context and Depth-aware Features for 3D Vehicle Detection from Point Cloud. ICRA 2020.
186	PointRGCN			92.15 %	97.48 %	86.83 %	0.26 s	GPU @ V100 (Python)
J. Zarzar, S. Giancola and B. Ghanem: PointRGCN: Graph Convolution Networks for 3D Vehicles Detection Refinement. ArXiv 2019.
187	EAEPNet			92.01 %	96.09 %	86.92 %	0.1 s	1 core @ 2.5 Ghz (Python)

188	F-ConvNet		code	91.98 %	95.81 %	79.83 %	0.47 s	GPU @ 2.5 Ghz (Python + C/C++)
Z. Wang and K. Jia: Frustum ConvNet: Sliding Frustums to Aggregate Local Point-Wise Features for Amodal 3D Object Detection. IROS 2019.
189	AB3DMOT		code	91.87 %	95.86 %	86.78 %	0.0047s	1 core @ 2.5 Ghz (Python)
X. Weng and K. Kitani: A Baseline for 3D Multi-Object Tracking. arXiv:1907.03961 2019.
190	geo-pillars			91.87 %	95.27 %	88.73 %	0.05 s	1 core @ 2.5 Ghz (C/C++)
ERROR: Wrong syntax in BIBTEX file.
191	SecAtten			91.83 %	95.03 %	88.54 %	0.1 s	1 core @ 2.5 Ghz (C/C++)

192	PASS-PointPillar			91.82 %	95.15 %	88.31 %	1 s	1 core @ 2.5 Ghz (C/C++)
Anonymous: Leveraging Anchor-based LiDAR 3D Object Detection via Point Assisted Sample Selection. will submit to computer vision conference/journal 2024.
193	MMLab-PointRCNN		code	91.77 %	95.90 %	86.92 %	0.1 s	GPU @ 2.5 Ghz (Python + C/C++)
S. Shi, X. Wang and H. Li: Pointrcnn: 3d object proposal generation and detection from point cloud. Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition 2019.
194	MMLab-PartA^2		code	91.73 %	95.00 %	88.86 %	0.08 s	GPU @ 2.5 Ghz (Python + C/C++)
S. Shi, Z. Wang, J. Shi, X. Wang and H. Li: From Points to Parts: 3D Object Detection from Point Cloud with Part-aware and Part-aggregation Network. IEEE Transactions on Pattern Analysis and Machine Intelligence 2020.
195	PCNet3D++			91.72 %	94.93 %	88.32 %	0.5 s	GPU @ 3.5 Ghz (Python)

196	DensePointPillars			91.70 %	95.37 %	86.38 %	0.03 s	GPU @ 2.5 Ghz (Python)

197	BFT3D_easy			91.61 %	97.11 %	84.18 %	0.18 s	1 core @ 2.5 Ghz (C/C++)

198	C-GCN			91.57 %	95.63 %	86.13 %	0.147 s	GPU @ V100 (Python)
J. Zarzar, S. Giancola and B. Ghanem: PointRGCN: Graph Convolution Networks for 3D Vehicles Detection Refinement. ArXiv 2019.
199	PCNet3D			91.44 %	95.05 %	87.92 %	0.05 s	GPU @ 3.5 Ghz (Python)

200	PL++: PV-RCNN++			91.42 %	94.76 %	88.30 %	0.342 s	RTX 4060Ti (Python)

201	PointRGBNet			91.33 %	95.39 %	86.29 %	0.08 s	4 cores @ 2.5 Ghz (Python + C/C++)
P. Xie Desheng: Real-time Detection of 3D Objects Based on Multi-Sensor Information Fusion. Automotive Engineering 2022.
202	mmFUSION		code	91.30 %	95.47 %	86.33 %	1s	1 core @ 2.5 Ghz (Python)
J. Ahmad and A. Del Bue: mmFUSION: Multimodal Fusion for 3D Objects Detection. arXiv preprint arXiv:2311.04058 2023.
203	SeSame-pillar		code	91.26 %	95.07 %	87.94 %	N/A s	TITAN RTX @ 1.35 Ghz (Python)
H. O, C. Yang and K. Huh: SeSame: Simple, Easy 3D Object Detection with Point-Wise Semantics. Proceedings of the Asian Conference on Computer Vision (ACCV) 2024.
204	EgoNet		code	91.23 %	96.11 %	80.96 %	0.1 s	GPU @ 1.5 Ghz (Python)
S. Li, Z. Yan, H. Li and K. Cheng: Exploring intermediate representation for monocular vehicle pose estimation. The IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) 2021.
205	Anonymous			91.13 %	94.00 %	87.90 %	0.02 s	GPU @ 2.5 Ghz (Python)

206	MonoDTF			91.13 %	94.92 %	85.52 %	0.1 s	1 core @ 2.5 Ghz (Python)

207	PFF3D		code	91.06 %	94.86 %	86.28 %	0.05 s	GPU @ 3.0 Ghz (Python + C/C++)
L. Wen and K. Jo: Fast and Accurate 3D Object Detection for Lidar-Camera-Based Autonomous Vehicles Using One Shared Voxel-Based Backbone. IEEE Access 2021.
208	SeSame-pillar w/scor		code	91.03 %	94.83 %	87.65 %	N/A s	1 core @ 2.5 Ghz (C/C++)
H. O, C. Yang and K. Huh: SeSame: Simple, Easy 3D Object Detection with Point-Wise Semantics. Proceedings of the Asian Conference on Computer Vision (ACCV) 2024.
209	zqd_test2			91.03 %	94.89 %	85.22 %	0.5 s	1 core @ 2.5 Ghz (C/C++)

210	Stereo CenterNet			91.02 %	96.54 %	83.15 %	0.04 s	GPU @ 2.5 Ghz (Python)
Y. Shi, Y. Guo, Z. Mi and X. Li: Stereo CenterNet-based 3D object detection for autonomous driving. Neurocomputing 2022.
211	MonoSKD		code	90.96 %	96.58 %	83.10 %	0.04 s	1 core @ 2.5 Ghz (Python)
S. Wang and J. Zheng: MonoSKD: General Distillation Framework for Monocular 3D Object Detection via Spearman Correlation Coefficient. ECAI 2023.
212	zqd			90.96 %	94.79 %	85.07 %	0.1 s	1 core @ 2.5 Ghz (Python)

213	AEPF			90.93 %	95.07 %	86.03 %	0.05 s	GPU @ 2.5 Ghz (Python)

214	Mix-Teaching		code	90.84 %	96.31 %	83.11 %	30 s	1 core @ 2.5 Ghz (C/C++)
L. Yang, X. Zhang, L. Wang, M. Zhu, C. Zhang and J. Li: Mix-Teaching: A Simple, Unified and Effective Semi-Supervised Learning Framework for Monocular 3D Object Detection. ArXiv 2022.
215	MonoFlex			90.82 %	95.95 %	83.11 %	0.03 s	GPU @ 2.5 Ghz (Python)
Y. Zhang, J. Lu and J. Zhou: Objects are Different: Flexible Monocular 3D Object Detection. CVPR 2021.
216	MonoCoP-Car			90.79 %	93.49 %	81.28 %	0.01 s	GPU @ 2.5 Ghz (Python)

217	Harmonic PointPillar		code	90.78 %	94.23 %	87.42 %	0.01 s	1 core @ 2.5 Ghz (Python)
H. Zhang, J. Mekala, Z. Nain, J. Park and H. Jung: 3D Harmonic Loss: Towards Task-consistent and Time-friendly 3D Object Detection for V2X Orchestration. will submit to IEEE Transactions on Vehicular Technology 2022.
218	MAFF-Net(DAF-Pillar)			90.78 %	94.17 %	83.17 %	0.04 s	1 core @ 2.5 Ghz (Python + C/C++)
Z. Zhang, Z. Liang, M. Zhang, X. Zhao, Y. Ming, T. Wenming and S. Pu: MAFF-Net: Filter False Positive for 3D Vehicle Detection with Multi-modal Adaptive Feature Fusion. arXiv preprint arXiv:2009.10945 2020.
219	HRI-VoxelFPN			90.76 %	96.35 %	85.37 %	0.02 s	GPU @ 2.5 Ghz (Python + C/C++)
H. Kuang, B. Wang, J. An, M. Zhang and Z. Zhang: Voxel-FPN:multi-scale voxel feature aggregation in 3D object detection from point clouds. sensors 2020.
220	KM3D		code	90.70 %	96.34 %	80.72 %	0.03 s	1 core @ 2.5 Ghz (Python)
P. Li: Monocular 3D Detection with Geometric Constraints Embedding and Semi-supervised Training. 2020.
221	PointPillars		code	90.70 %	93.84 %	87.47 %	16 ms	1080ti GPU and Intel i7 CPU
A. Lang, S. Vora, H. Caesar, L. Zhou, J. Yang and O. Beijbom: PointPillars: Fast Encoders for Object Detection from Point Clouds. CVPR 2019.
222	WS3D			90.69 %	94.85 %	85.94 %	0.1 s	GPU @ 2.5 Ghz (Python)
Q. Meng, W. Wang, T. Zhou, J. Shen, L. Van Gool and D. Dai: Weakly Supervised 3D Object Detection from Lidar Point Cloud. 2020.
223	EOTL		code	90.67 %	96.14 %	80.59 %	TBD s	1 core @ 2.5 Ghz (Python + C/C++)
R. Yang, Z. Yan, T. Yang, Y. Wang and Y. Ruichek: Efficient Online Transfer Learning for Road Participants Detection in Autonomous Driving. IEEE Sensors Journal 2023.
224	DCD		code	90.66 %	96.31 %	83.01 %	0.03 s	1 core @ 2.5 Ghz (C/C++)
Y. Li, Y. Chen, J. He and Z. Zhang: Densely Constrained Depth Estimator for Monocular 3D Object Detection. European Conference on Computer Vision 2022.
225	NeurOCS			90.66 %	96.15 %	80.64 %	0.1 s	GPU @ 2.5 Ghz (Python)
Z. Min, B. Zhuang, S. Schulter, B. Liu, E. Dunn and M. Chandraker: NeurOCS: Neural NOCS Supervision for Monocular 3D Object Localization. CVPR 2023.
226	MonoEF			90.65 %	96.19 %	82.95 %	0.03 s	1 core @ 2.5 Ghz (Python)
Y. Zhou, Y. He, H. Zhu, C. Wang, H. Li and Q. Jiang: Monocular 3D Object Detection: An Extrinsic Parameter Free Approach. Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) 2021.
227	CIE			90.64 %	96.19 %	82.90 %	0.1 s	1 core @ 2.5 Ghz (C/C++)
Anonymities: Consistency of Implicit and Explicit Features Matters for Monocular 3D Object Detection. arXiv preprint arXiv:2207.07933 2022.
228	T-SSD			90.56 %	96.05 %	85.37 %	0.04	1 core @ 2.0 Ghz (C/C++)

229	DID-M3D		code	90.55 %	94.20 %	80.61 %	0.04 s	1 core @ 2.5 Ghz (Python)
L. Peng, X. Wu, Z. Yang, H. Liu and D. Cai: DID-M3D: Decoupling Instance Depth for Monocular 3D Object Detection. ECCV 2022.
230	QD-3DT		code	90.49 %	92.61 %	80.32 %	0.03 s	GPU @ 2.5 Ghz (Python)
H. Hu, Y. Yang, T. Fischer, F. Yu, T. Darrell and M. Sun: Monocular Quasi-Dense 3D Object Tracking. ArXiv:2103.07351 2021.
231	HomoLoss(monoflex)		code	90.49 %	95.86 %	80.66 %	0.04 s	1 core @ 2.5 Ghz (Python)
J. Gu, B. Wu, L. Fan, J. Huang, S. Cao, Z. Xiang and X. Hua: Homography Loss for Monocular 3D Object Detection. Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) 2022.
232	SeSame-voxel		code	90.42 %	95.76 %	87.40 %	N/A s	TITAN RTX @ 1.35 Ghz (Python)
H. O, C. Yang and K. Huh: SeSame: Simple, Easy 3D Object Detection with Point-Wise Semantics. Proceedings of the Asian Conference on Computer Vision (ACCV) 2024.
233	MonoQ			90.36 %	96.03 %	80.29 %	0.02 s	1 core @ 2.5 Ghz (Python)

234	monodle		code	90.23 %	93.46 %	80.11 %	0.04 s	GPU @ 2.5 Ghz (Python)
X. Ma, Y. Zhang, D. Xu, D. Zhou, S. Yi, H. Li and W. Ouyang: Delving into Localization Errors for Monocular 3D Object Detection. CVPR 2021 .
235	3D IoU Loss			90.21 %	95.60 %	84.96 %	0.08 s	GPU @ 2.5 Ghz (Python + C/C++)
D. Zhou, J. Fang, X. Song, C. Guan, J. Yin, Y. Dai and R. Yang: IoU Loss for 2D/3D Object Detection. International Conference on 3D Vision (3DV) 2019.
236	MonoCInIS			90.20 %	95.80 %	82.00 %	0,13 s	GPU @ 2.5 Ghz (C/C++)
J. Heylen, M. De Wolf, B. Dawagne, M. Proesmans, L. Van Gool, W. Abbeloos, H. Abdelkawy and D. Reino: MonoCInIS: Camera Independent Monocular 3D Object Detection using Instance Segmentation. Proceedings of the IEEE/CVF International Conference on Computer Vision 2021.
237	ARPNET			90.11 %	93.42 %	82.56 %	0.08 s	GPU @ 2.5 Ghz (Python + C/C++)
Y. Ye, C. Zhang and X. Hao: ARPNET: attention region proposal network for 3D object detection. Science China Information Sciences 2019.
238	TANet		code	90.11 %	93.52 %	84.61 %	0.035s	GPU @ 2.5 Ghz (Python + C/C++)
Z. Liu, X. Zhao, T. Huang, R. Hu, Y. Zhou and X. Bai: TANet: Robust 3D Object Detection from Point Clouds with Triple Attention. AAAI 2020.
239	CG-Stereo			89.98 %	96.28 %	82.21 %	0.57 s	GeForce RTX 2080 Ti
C. Li, J. Ku and S. Waslander: Confidence Guided Stereo 3D Object Detection with Split Depth Estimation. IROS 2020.
240	Deep3DBox			89.88 %	94.62 %	76.40 %	1.5 s	GPU @ 2.5 Ghz (C/C++)
A. Mousavian, D. Anguelov, J. Flynn and J. Kosecka: 3D Bounding Box Estimation Using Deep Learning and Geometry. CVPR 2017.
241	CMKD		code	89.81 %	95.07 %	83.24 %	0.1 s	1 core @ 2.5 Ghz (C/C++)
Y. Hong, H. Dai and Y. Ding: Cross-Modality Knowledge Distillation Network for Monocular 3D Object Detection. ECCV 2022.
242	PS-fld		code	89.78 %	95.60 %	81.68 %	0.25 s	1 core @ 2.5 Ghz (C/C++)
Y. Chen, H. Dai and Y. Ding: Pseudo-Stereo for Monocular 3D Object Detection in Autonomous Driving. Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) 2022.
243	GPP		code	89.68 %	93.94 %	80.60 %	0.23 s	GPU @ 1.5 Ghz (Python + C/C++)
A. Rangesh and M. Trivedi: Ground plane polling for 6dof pose estimation of objects on the road. IEEE Transactions on Intelligent Vehicles 2020.
244	SubCNN			89.53 %	94.11 %	79.14 %	2 s	GPU @ 3.5 Ghz (Python + C/C++)
Y. Xiang, W. Choi, Y. Lin and S. Savarese: Subcategory-aware Convolutional Neural Networks for Object Proposals and Detection. IEEE Winter Conference on Applications of Computer Vision (WACV) 2017.
245	HINTED		code	89.41 %	93.97 %	83.95 %	0.04 s	1 core @ 2.5 Ghz (C/C++)
Q. Xia, W. Ye, H. Wu, S. Zhao, L. Xing, X. Huang, J. Deng, X. Li, C. Wen and C. Wang: HINTED: Hard Instance Enhanced Detector with Mixed-Density Feature Fusion for Sparsely- Supervised 3D Object Detection. Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition 2024.
246	SCNet			89.36 %	95.23 %	84.03 %	0.04 s	GPU @ 3.0 Ghz (Python)
Z. Wang, H. Fu, L. Wang, L. Xiao and B. Dai: SCNet: Subdivision Coding Network for Object Detection Based on 3D Point Cloud. IEEE Access 2019.
247	AVOD		code	89.22 %	94.98 %	82.14 %	0.08 s	Titan X (pascal)
J. Ku, M. Mozifian, J. Lee, A. Harakeh and S. Waslander: Joint 3D Proposal Generation and Object Detection from View Aggregation. IROS 2018.
248	MonoCoP			89.15 %	91.77 %	79.47 %	0.01 s	1 core @ 2.5 Ghz (C/C++)

249	IAFA			89.14 %	92.96 %	79.40 %	0.04 s	1 core @ 2.5 Ghz (C/C++)
D. Zhou, X. Song, Y. Dai, J. Yin, F. Lu, M. Liao, J. Fang and L. Zhang: IAFA: Instance-Aware Feature Aggregation for 3D Object Detection from a Single Image. Proceedings of the Asian Conference on Computer Vision 2020.
250	MonoDDE			89.07 %	96.72 %	81.42 %	0.04 s	1 core @ 2.5 Ghz (Python)

	Method	Setting	Code	Moderate	Easy	Hard	Runtime	Environment
1	VMVS			68.19 %	79.98 %	63.18 %	0.25 s	GPU @ 2.5 Ghz (Python)
J. Ku, A. Pon, S. Walsh and S. Waslander: Improving 3D object detection for pedestrians with virtual multi-view synthesis orientation estimation. IROS 2019.
2	SubCNN			66.70 %	79.65 %	61.35 %	2 s	GPU @ 3.5 Ghz (Python + C/C++)
Y. Xiang, W. Choi, Y. Lin and S. Savarese: Subcategory-aware Convolutional Neural Networks for Object Proposals and Detection. IEEE Winter Conference on Applications of Computer Vision (WACV) 2017.
3	DD3D		code	63.92 %	77.09 %	59.41 %	n/a s	1 core @ 2.5 Ghz (C/C++)
D. Park, R. Ambrus, V. Guizilini, J. Li and A. Gaidon: Is Pseudo-Lidar needed for Monocular 3D Object detection?. IEEE/CVF International Conference on Computer Vision (ICCV) .
4	F-ConvNet		code	63.87 %	75.19 %	58.57 %	0.47 s	GPU @ 2.5 Ghz (Python + C/C++)
Z. Wang and K. Jia: Frustum ConvNet: Sliding Frustums to Aggregate Local Point-Wise Features for Amodal 3D Object Detection. IROS 2019.
5	UPIDet		code	61.92 %	72.38 %	59.31 %	0.11 s	1 core @ 2.5 Ghz (C/C++)
Y. Zhang, Q. Zhang, J. Hou, Y. Yuan and G. Xing: Unleash the Potential of Image Branch for Cross-modal 3D Object Detection. Thirty-seventh Conference on Neural Information Processing Systems 2023.
6	LCANet			61.63 %	72.07 %	57.48 %	1 s	1 core @ 2.5 Ghz (C/C++)

7	CasA++		code	61.59 %	71.78 %	58.71 %	0.1 s	1 core @ 2.5 Ghz (C/C++)
H. Wu, J. Deng, C. Wen, X. Li and C. Wang: CasA: A Cascade Attention Network for 3D Object Detection from LiDAR point clouds. IEEE Transactions on Geoscience and Remote Sensing 2022.
8	3DOP		code	61.48 %	74.22 %	55.89 %	3s	GPU @ 2.5 Ghz (Matlab + C/C++)
X. Chen, K. Kundu, Y. Zhu, A. Berneshawi, H. Ma, S. Fidler and R. Urtasun: 3D Object Proposals for Accurate Object Class Detection. NIPS 2015.
9	TED		code	61.44 %	71.72 %	58.59 %	0.1 s	1 core @ 2.5 Ghz (C/C++)
H. Wu, C. Wen, W. Li, R. Yang and C. Wang: Transformation-Equivariant 3D Object Detection for Autonomous Driving. AAAI 2023.
10	LoGoNet		code	60.70 %	69.16 %	58.13 %	0.1 s	1 core @ 2.5 Ghz (C/C++)
X. Li, T. Ma, Y. Hou, B. Shi, Y. Yang, Y. Liu, X. Wu, Q. Chen, Y. Li, Y. Qiao and others: LoGoNet: Towards Accurate 3D Object Detection with Local-to-Global Cross-Modal Fusion. CVPR 2023.
11	HotSpotNet			60.65 %	70.36 %	57.42 %	0.04 s	1 core @ 2.5 Ghz (Python + C/C++)
Q. Chen, L. Sun, Z. Wang, K. Jia and A. Yuille: object as hotspots. Proceedings of the European Conference on Computer Vision (ECCV) 2020.
12	MonoAFKD			60.39 %	75.23 %	53.94 %	0.03 s	1 core @ 2.5 Ghz (C/C++)

13	MonoLSS			60.28 %	75.13 %	53.85 %	0.04 s	1 core @ 2.5 Ghz (Python)
Z. Li, J. Jia and Y. Shi: MonoLSS: Learnable Sample Selection For Monocular 3D Detection. International Conference on 3D Vision 2024.
14	DeepStereoOP			60.15 %	73.76 %	55.30 %	3.4 s	GPU @ 3.5 Ghz (Matlab + C/C++)
C. Pham and J. Jeon: Robust Object Proposals Re-ranking for Object Detection in Autonomous Driving Using Convolutional Neural Networks. Signal Processing: Image Communiation 2017.
15	Pose-RCNN			59.84 %	76.24 %	53.59 %	2 s	>8 cores @ 2.5 Ghz (Python)
M. Braun, Q. Rao, Y. Wang and F. Flohr: Pose-RCNN: Joint object detection and pose estimation using 3D object proposals. Intelligent Transportation Systems (ITSC), 2016 IEEE 19th International Conference on 2016.
16	USVLab BSAODet		code	59.73 %	69.95 %	55.85 %	0.04 s	1 core @ 2.5 Ghz (C/C++)
W. Xiao, Y. Peng, C. Liu, J. Gao, Y. Wu and X. Li: Balanced Sample Assignment and Objective for Single-Model Multi-Class 3D Object Detection. IEEE Transactions on Circuits and Systems for Video Technology 2023.
17	CasA		code	59.69 %	70.33 %	56.89 %	0.1 s	1 core @ 2.5 Ghz (C/C++)
H. Wu, J. Deng, C. Wen, X. Li and C. Wang: CasA: A Cascade Attention Network for 3D Object Detection from LiDAR point clouds. IEEE Transactions on Geoscience and Remote Sensing 2022.
18	FFNet		code	58.87 %	69.24 %	53.75 %	1.07 s	GPU @ 1.5 Ghz (Python)
C. Zhao, Y. Qian and M. Yang: Monocular Pedestrian Orientation Estimation Based on Deep 2D-3D Feedforward. Pattern Recognition 2019.
19	Mono3D		code	58.66 %	71.19 %	53.94 %	4.2 s	GPU @ 2.5 Ghz (Matlab + C/C++)
X. Chen, K. Kundu, Z. Zhang, H. Ma, S. Fidler and R. Urtasun: Monocular 3D Object Detection for Autonomous Driving. CVPR 2016.
20	VPFNet		code	58.63 %	67.96 %	54.77 %	0.2 s	1 core @ 2.5 Ghz (C/C++)
C. Wang, H. Chen and L. Fu: VPFNet: Voxel-Pixel Fusion Network for Multi-class 3D Object Detection. 2021. C. Wang, H. Chen, Y. Chen, P. Hsiao and L. Fu: VoPiFNet: Voxel-Pixel Fusion Network for Multi-Class 3D Object Detection. IEEE Transactions on Intelligent Transportation Systems 2024.
21	P2V-RCNN			57.94 %	68.67 %	55.07 %	0.1 s	2 cores @ 2.5 Ghz (Python)
J. Li, S. Luo, Z. Zhu, H. Dai, A. Krylov, Y. Ding and L. Shao: P2V-RCNN: Point to Voxel Feature Learning for 3D Object Detection from Point Clouds. IEEE Access 2021.
22	OGMMDet		code	57.91 %	69.64 %	55.24 %	0.01 s	1 core @ 2.5 Ghz (C/C++)

23	ANM		code	57.91 %	69.64 %	55.24 %	ANM	ANM

24	dsvd+vx			57.68 %	68.29 %	54.96 %	0.1 s	1 core @ 2.5 Ghz (C/C++)

25	Fast-CLOCs			57.35 %	70.93 %	54.48 %	0.1 s	GPU @ 2.5 Ghz (Python)
S. Pang, D. Morris and H. Radha: Fast-CLOCs: Fast Camera-LiDAR Object Candidates Fusion for 3D Object Detection. Proceedings of the IEEE/CVF Winter Conference on Applications of Computer Vision (WACV) 2022.
26	EOTL		code	57.17 %	68.99 %	51.48 %	TBD s	1 core @ 2.5 Ghz (Python + C/C++)
R. Yang, Z. Yan, T. Yang, Y. Wang and Y. Ruichek: Efficient Online Transfer Learning for Road Participants Detection in Autonomous Driving. IEEE Sensors Journal 2023.
27	MLF-DET			56.89 %	64.49 %	53.17 %	0.09 s	1 core @ 2.5 Ghz (C/C++)
Z. Lin, Y. Shen, S. Zhou, S. Chen and N. Zheng: MLF-DET: Multi-Level Fusion for Cross- Modal 3D Object Detection. International Conference on Artificial Neural Networks 2023.
28	FIRM-Net_SCF+			56.65 %	66.97 %	53.90 %	0.07 s	1 core @ 2.5 Ghz (Python)

29	LPFusion_three_class			56.46 %	66.96 %	52.51 %	0.1 s	1 core @ 2.5 Ghz (C/C++)

30	FIRM-Net-SCF			56.44 %	66.80 %	53.65 %	0.07 s	1 core @ 2.5 Ghz (Python)

31	BVIFusion+			56.36 %	66.56 %	52.55 %	0.09 s	1 core @ 2.5 Ghz (Python)

32	SCNet3D			55.83 %	65.98 %	53.07 %	0.08 s	1 core @ 2.5 Ghz (Python)

33	LumiNet		code	55.80 %	66.85 %	53.17 %	0.1 s	1 core @ 2.5 Ghz (Python)

34	AFFN-G			55.17 %	63.53 %	52.79 %	0.5 s	1 core @ 2.5 Ghz (C/C++)

35	GEFPN			55.17 %	63.53 %	52.79 %	0.5 s	1 core @ 2.5 Ghz (C/C++)

36	GeVo			55.17 %	63.53 %	52.79 %	0.05 s	1 core @ 2.5 Ghz (C/C++)

37	vsis-PHNet			55.13 %	66.56 %	52.44 %	0.03 s	1 core @ 2.5 Ghz (C/C++)

38	PHNetp			55.13 %	66.56 %	52.44 %	0.03 s	1 core @ 2.5 Ghz (C/C++)

39	DFAF3D			54.99 %	65.42 %	51.21 %	0.05 s	1 core @ 2.5 Ghz (Python)
Q. Tang, X. Bai, J. Guo, B. Pan and W. Jiang: DFAF3D: A dual-feature-aware anchor-free single-stage 3D detector for point clouds. Image and Vision Computing 2023.
40	MPC3DNet			54.92 %	63.60 %	52.36 %	0.05 s	GPU @ 1.5 Ghz (Python)

41	3ONet			54.88 %	66.35 %	50.82 %	0.1 s	1 core @ 2.5 Ghz (C/C++)
H. Hoang and M. Yoo: 3ONet: 3-D Detector for Occluded Object Under Obstructed Conditions. IEEE Sensors Journal 2023.
42	FromVoxelToPoint		code	54.80 %	66.21 %	52.03 %	0.1 s	1 core @ 2.5 Ghz (C/C++)
J. Li, H. Dai, L. Shao and Y. Ding: From Voxel to Point: IoU-guided 3D Object Detection for Point Cloud with Voxel-to- Point Decoder. MM '21: The 29th ACM International Conference on Multimedia (ACM MM) 2021.
43	test			54.79 %	63.76 %	52.50 %	0.04 s	GPU @ 1.5 Ghz (Python + C/C++)

44	MonoPSR		code	54.65 %	68.98 %	50.07 %	0.2 s	GPU @ 3.5 Ghz (Python)
J. Ku, A. Pon and S. Waslander: Monocular 3D Object Detection Leveraging Accurate Proposals and Shape Reconstruction. CVPR 2019.
45	SVFMamba		code	54.57 %	65.47 %	50.98 %	N/A s	1 core @ 2.5 Ghz (C/C++)

46	DSA-PV-RCNN		code	54.38 %	63.12 %	51.98 %	0.08 s	1 core @ 2.5 Ghz (Python + C/C++)
P. Bhattacharyya, C. Huang and K. Czarnecki: SA-Det3D: Self-Attention Based Context-Aware 3D Object Detection. 2021.
47	PDV		code	54.08 %	63.43 %	50.75 %	0.1 s	1 core @ 2.5 Ghz (C/C++)
J. Hu, T. Kuai and S. Waslander: Point Density-Aware Voxels for LiDAR 3D Object Detection. CVPR 2022.
48	ACFNet			53.97 %	65.55 %	49.97 %	0.11 s	1 core @ 2.5 Ghz (C/C++)
Y. Tian, X. Zhang, X. Wang, J. Xu, J. Wang, R. Ai, W. Gu and W. Ding: ACF-Net: Asymmetric Cascade Fusion for 3D Detection With LiDAR Point Clouds and Images. IEEE Transactions on Intelligent Vehicles 2023.
49	PASS-PV-RCNN-Plus			53.82 %	63.49 %	51.30 %	1 s	1 core @ 2.5 Ghz (Python)
Anonymous: Leveraging Anchor-based LiDAR 3D Object Detection via Point Assisted Sample Selection. will submit to computer vision conference/journal 2024.
50	monodle		code	53.78 %	69.94 %	48.98 %	0.04 s	GPU @ 2.5 Ghz (Python)
X. Ma, Y. Zhang, D. Xu, D. Zhou, S. Yi, H. Li and W. Ouyang: Delving into Localization Errors for Monocular 3D Object Detection. CVPR 2021 .
51	AFFN			53.77 %	63.18 %	50.98 %	0.5 s	1 core @ 2.5 Ghz (C/C++)

52	MGAF-3DSSD		code	53.73 %	64.69 %	49.84 %	0.1 s	1 core @ 2.5 Ghz (C/C++)
J. Li, H. Dai, L. Shao and Y. Ding: Anchor-free 3D Single Stage Detector with Mask-Guided Attention for Point Cloud. MM '21: The 29th ACM International Conference on Multimedia (ACM MM) 2021.
53	HMNet			53.48 %	65.59 %	50.71 %	0.1 s	1 core @ 2.5 Ghz (C/C++)

54	MLFusion-VS			53.46 %	62.55 %	50.98 %	0.06 s	1 core @ 2.5 Ghz (C/C++)

55	SFA-GCL(80, k=4)		code	53.43 %	66.21 %	50.74 %	0.04 s	1 core @ 2.5 Ghz (C/C++)

56	AM			53.39 %	71.12 %	48.57 %	0.04 s	1 core @ 2.5 Ghz (C/C++)

57	SRDL			53.36 %	63.39 %	50.43 %	0.05 s	1 core @ 2.5 Ghz (Python + C/C++)
ERROR: Wrong syntax in BIBTEX file.
58	SFA-GCL(80)		code	53.26 %	64.46 %	48.99 %	0.04 s	1 core @ 2.5 Ghz (C/C++)

59	PG-RCNN		code	53.12 %	63.73 %	50.46 %	0.06 s	GPU @ 1.5 Ghz (Python)
I. Koo, I. Lee, S. Kim, H. Kim, W. Jeon and C. Kim: PG-RCNN: Semantic Surface Point Generation for 3D Object Detection. 2023.
60	SFA-GCL_dataaug		code	53.10 %	65.98 %	50.38 %	0.04 s	1 core @ 2.5 Ghz (C/C++)

61	SFA-GCL		code	52.98 %	65.70 %	50.30 %	0.04 s	1 core @ 2.5 Ghz (C/C++)

62	SFA-GCL(baseline)		code	52.83 %	65.53 %	50.14 %	0.04 s	1 core @ 2.5 Ghz (C/C++)

63	IA-SSD (single)		code	52.69 %	62.90 %	50.27 %	0.013 s	1 core @ 2.5 Ghz (C/C++)
Y. Zhang, Q. Hu, G. Xu, Y. Ma, J. Wan and Y. Guo: Not All Points Are Equal: Learning Highly Efficient Point-based Detectors for 3D LiDAR Point Clouds. CVPR 2022.
64	R2Pfusion-Det			52.67 %	63.49 %	50.39 %	0.3 s	1 core @ 2.5 Ghz (C/C++)

65	Voxel RCNN-Focal*		code	52.65 %	60.83 %	50.40 %	0.2 s	1 core @ 2.5 Ghz (Python)

66	MonoUNI		code	52.62 %	69.15 %	47.89 %	0.04 s	1 core @ 2.5 Ghz (Python)
J. Jia, Z. Li and Y. Shi: MonoUNI: A Unified Vehicle and Infrastructure-side Monocular 3D Object Detection Network with Sufficient Depth Clues. Thirty-seventh Conference on Neural Information Processing Systems 2023.
67	CG-SSD			52.59 %	62.52 %	49.89 %	0.01 s	1 core @ 2.5 Ghz (C/C++)

68	HMFI		code	52.47 %	63.10 %	49.57 %	0.1 s	1 core @ 2.5 Ghz (C/C++)
X. Li, B. Shi, Y. Hou, X. Wu, T. Ma, Y. Li and L. He: Homogeneous Multi-modal Feature Fusion and Interaction for 3D Object Detection. ECCV 2022.
69	MMLab PV-RCNN		code	52.42 %	63.45 %	49.23 %	0.08 s	1 core @ 2.5 Ghz (Python + C/C++)
S. Shi, C. Guo, L. Jiang, Z. Wang, J. Shi, X. Wang and H. Li: PV-RCNN: Point-Voxel Feature Set Abstraction for 3D Object Detection. CVPR 2020.
70	SVGA-Net			52.27 %	62.33 %	49.44 %	0.03s	1 core @ 2.5 Ghz (Python + C/C++)
Q. He, Z. Wang, H. Zeng, Y. Zeng and Y. Liu: SVGA-Net: Sparse Voxel-Graph Attention Network for 3D Object Detection from Point Clouds. AAAI 2022.
71	MMLab-PartA^2		code	52.20 %	63.51 %	48.27 %	0.08 s	GPU @ 2.5 Ghz (Python + C/C++)
S. Shi, Z. Wang, J. Shi, X. Wang and H. Li: From Points to Parts: 3D Object Detection from Point Cloud with Part-aware and Part-aggregation Network. IEEE Transactions on Pattern Analysis and Machine Intelligence 2020.
72	FRCNN+Or		code	52.15 %	67.03 %	47.14 %	0.09 s	Titan Xp GPU
C. Guindel, D. Martin and J. Armingol: Fast Joint Object Detection and Viewpoint Estimation for Traffic Scene Understanding. IEEE Intelligent Transportation Systems Magazine 2018. C. Guindel, D. Martin and J. Armingol: Joint Object Detection and Viewpoint Estimation using CNN features. IEEE International Conference on Vehicular Electronics and Safety (ICVES) 2017.
73	CAIA_PRO		code	52.11 %	62.38 %	49.56 %	0.01 s	1 core @ 2.5 Ghz (C/C++)

74	SIF			52.10 %	62.72 %	49.19 %	0.1 s	1 core @ 2.5 Ghz (C/C++)
P. An: SIF. Submitted to CVIU 2021.
75	FocalsConv*			51.97 %	59.61 %	48.59 %	0.2 s	1 core @ 2.5 Ghz (C/C++)

76	CGML			51.91 %	60.01 %	49.53 %	0.33 s	1 core @ 2.5 Ghz (C/C++)

77	R50_SACINet			51.91 %	62.91 %	48.32 %	0.06 s	1 core @ 2.5 Ghz (C/C++)

78	SFA_IGCL_Focalsconv*		code	51.76 %	60.17 %	49.28 %	0.2 s	1 core @ 2.5 Ghz (C/C++)

79	AFFN-Ga			51.63 %	60.97 %	49.10 %	0.5 s	1 core @ 2.5 Ghz (C/C++)

80	QD-3DT		code	51.46 %	68.64 %	47.00 %	0.03 s	GPU @ 2.5 Ghz (Python)
H. Hu, Y. Yang, T. Fischer, F. Yu, T. Darrell and M. Sun: Monocular Quasi-Dense 3D Object Tracking. ArXiv:2103.07351 2021.
81	MFB3D			51.36 %	59.89 %	48.44 %	0.14 s	1 core @ 2.5 Ghz (Python)

82	HA-PillarNet			51.11 %	58.89 %	48.91 %	0.05 s	1 core @ 2.5 Ghz (Python)

83	SFA-GCL		code	51.00 %	63.63 %	46.80 %	0.04 s	1 core @ 2.5 Ghz (C/C++)

84	ACDet		code	50.90 %	62.39 %	48.34 %	0.05 s	1 core @ 2.5 Ghz (C/C++)
J. Xu, G. Wang, X. Zhang and G. Wan: ACDet: Attentive Cross-view Fusion for LiDAR-based 3D Object Detection. 3DV 2022.
85	MonoMH		code	50.76 %	68.08 %	45.95 %	0.04 s	1 core @ 2.5 Ghz (Python)

86	GUPNet		code	50.74 %	68.93 %	44.01 %	NA s	1 core @ 2.5 Ghz (Python + C/C++)
Y. Lu, X. Ma, L. Yang, T. Zhang, Y. Liu, Q. Chu, J. Yan and W. Ouyang: Geometry Uncertainty Projection Network for Monocular 3D Object Detection. arXiv preprint arXiv:2107.13774 2021.
87	DEVIANT		code	50.66 %	68.78 %	45.89 %	0.04 s	1 GPU (Python)
A. Kumar, G. Brazil, E. Corona, A. Parchami and X. Liu: DEVIANT: Depth EquiVarIAnt NeTwork for Monocular 3D Object Detection. European Conference on Computer Vision (ECCV) 2022.
88	DVFENet			50.52 %	60.32 %	47.92 %	0.05 s	1 core @ 2.5 Ghz (Python + C/C++)
Y. He, G. Xia, Y. Luo, L. Su, Z. Zhang, W. Li and P. Wang: DVFENet: Dual-branch Voxel Feature Extraction Network for 3D Object Detection. Neurocomputing 2021.
89	OPA-3D		code	50.42 %	68.35 %	43.91 %	0.04 s	1 core @ 3.5 Ghz (Python)
Y. Su, Y. Di, G. Zhai, F. Manhardt, J. Rambach, B. Busam, D. Stricker and F. Tombari: OPA-3D: Occlusion-Aware Pixel-Wise Aggregation for Monocular 3D Object Detection. IEEE Robotics and Automation Letters 2023.
90	PointPainting			50.22 %	59.25 %	46.95 %	0.4 s	GPU @ 2.5 Ghz (Python + C/C++)
S. Vora, A. Lang, B. Helou and O. Beijbom: PointPainting: Sequential Fusion for 3D Object Detection. CVPR 2020.
91	Mix-Teaching		code	50.19 %	64.04 %	44.37 %	30 s	1 core @ 2.5 Ghz (C/C++)
L. Yang, X. Zhang, L. Wang, M. Zhu, C. Zhang and J. Li: Mix-Teaching: A Simple, Unified and Effective Semi-Supervised Learning Framework for Monocular 3D Object Detection. ArXiv 2022.
92	M3DeTR		code	50.09 %	58.90 %	47.66 %	n/a s	GPU @ 1.0 Ghz (Python)
T. Guan, J. Wang, S. Lan, R. Chandra, Z. Wu, L. Davis and D. Manocha: M3DeTR: Multi-representation, Multi- scale, Mutual-relation 3D Object Detection with Transformers. 2021.
93	SA V1			50.06 %	60.55 %	47.36 %	0.5 s	GPU @ 2.5 Ghz (Python)

94	L_SACINet			49.63 %	59.81 %	46.15 %	0.04 s	1 core @ 2.5 Ghz (C/C++)

95	IA-SSD (multi)		code	49.58 %	62.51 %	47.17 %	0.014 s	1 core @ 2.5 Ghz (C/C++)
Y. Zhang, Q. Hu, G. Xu, Y. Ma, J. Wan and Y. Guo: Not All Points Are Equal: Learning Highly Efficient Point-based Detectors for 3D LiDAR Point Clouds. CVPR 2022.
96	GSG-FPS		code	49.58 %	58.99 %	47.18 %	0.01 s	1 core @ 2.5 Ghz (C/C++)

97	SAKD-MR-Res18			49.57 %	67.47 %	43.12 %	0.03 s	1 core @ 2.5 Ghz (C/C++)

98	MSMA V1			49.31 %	58.85 %	46.35 %	0.5 s	GPU @ 2.5 Ghz (Python)

99	XView			49.30 %	58.39 %	46.81 %	0.1 s	1 core @ 2.5 Ghz (C/C++)
L. Xie, G. Xu, D. Cai and X. He: X-view: Non-egocentric Multi-View 3D Object Detector. 2021.
100	Test_dif		code	48.72 %	58.85 %	46.44 %	0.01 s	1 core @ 2.5 Ghz (C/C++)

101	ARPNET			48.49 %	60.47 %	45.02 %	0.08 s	GPU @ 2.5 Ghz (Python + C/C++)
Y. Ye, C. Zhang and X. Hao: ARPNET: attention region proposal network for 3D object detection. Science China Information Sciences 2019.
102	MonoCoP			48.49 %	64.34 %	43.91 %	0.01 s	1 core @ 2.5 Ghz (C/C++)

103	DPPFA-Net			48.38 %	56.13 %	45.93 %	0.1 s	1 core @ 2.5 Ghz (Python)
J. Wang, X. Kong, H. Nishikawa, Q. Lian and H. Tomiyama: Dynamic Point-Pixel Feature Alignment for Multi-modal 3D Object Detection. IEEE Internet of Things Journal 2023.
104	PointPillars		code	48.05 %	57.47 %	45.40 %	16 ms	1080ti GPU and Intel i7 CPU
A. Lang, S. Vora, H. Caesar, L. Zhou, J. Yang and O. Beijbom: PointPillars: Fast Encoders for Object Detection from Point Clouds. CVPR 2019.
105	HINTED		code	47.84 %	62.13 %	43.71 %	0.04 s	1 core @ 2.5 Ghz (C/C++)
Q. Xia, W. Ye, H. Wu, S. Zhao, L. Xing, X. Huang, J. Deng, X. Li, C. Wen and C. Wang: HINTED: Hard Instance Enhanced Detector with Mixed-Density Feature Fusion for Sparsely- Supervised 3D Object Detection. Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition 2024.
106	MonoRUn		code	47.82 %	63.28 %	43.23 %	0.07 s	GPU @ 2.5 Ghz (Python + C/C++)
H. Chen, Y. Huang, W. Tian, Z. Gao and L. Xiong: MonoRUn: Monocular 3D Object Detection by Reconstruction and Uncertainty Propagation. IEEE Conference on Computer Vision and Pattern Recognition (CVPR) 2021.
107	SeSame-voxel		code	47.60 %	58.80 %	43.53 %	N/A s	TITAN RTX @ 1.35 Ghz (Python)
H. O, C. Yang and K. Huh: SeSame: Simple, Easy 3D Object Detection with Point-Wise Semantics. Proceedings of the Asian Conference on Computer Vision (ACCV) 2024.
108	L-AUG			47.59 %	58.42 %	44.64 %	0.1 s	1 core @ 2.5 Ghz (C/C++)
T. Cortinhal, I. Gouigah and E. Aksoy: Semantics-aware LiDAR-Only Pseudo Point Cloud Generation for 3D Object Detection. 2023.
109	MMLab-PointRCNN		code	47.33 %	57.19 %	44.31 %	0.1 s	GPU @ 2.5 Ghz (Python + C/C++)
S. Shi, X. Wang and H. Li: Pointrcnn: 3d object proposal generation and detection from point cloud. Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition 2019.
110	Anonymous			47.25 %	58.69 %	44.49 %	0.02 s	GPU @ 2.5 Ghz (Python)

111	SeSame-point		code	47.09 %	56.55 %	44.58 %	N/A s	TITAN RTX @ 1.35 Ghz (Python)
H. O, C. Yang and K. Huh: SeSame: Simple, Easy 3D Object Detection with Point-Wise Semantics. Proceedings of the Asian Conference on Computer Vision (ACCV) 2024.
112	S-AT GCN			46.64 %	56.55 %	44.23 %	0.02 s	GPU @ 2.0 Ghz (Python)
L. Wang, C. Wang, X. Zhang, T. Lan and J. Li: S-AT GCN: Spatial-Attention Graph Convolution Network based Feature Enhancement for 3D Object Detection. CoRR 2021.
113	PiFeNet		code	46.59 %	55.11 %	44.14 %	0.03 s	1 core @ 2.5 Ghz (C/C++)
D. Le, H. Shi, H. Rezatofighi and J. Cai: Accurate and Real-time 3D Pedestrian Detection Using an Efficient Attentive Pillar Network. IEEE Robotics and Automation Letters 2022.
114	Shift R-CNN (mono)		code	46.56 %	64.73 %	41.86 %	0.25 s	GPU @ 1.5 Ghz (Python)
A. Naiden, V. Paunescu, G. Kim, B. Jeon and M. Leordeanu: Shift R-CNN: Deep Monocular 3D Object Detection With Closed-form Geometric Constraints. ICIP 2019.
115	Disp R-CNN		code	45.80 %	63.23 %	41.32 %	0.387 s	GPU @ 2.5 Ghz (Python + C/C++)
J. Sun, L. Chen, Y. Xie, S. Zhang, Q. Jiang, X. Zhou and H. Bao: Disp R-CNN: Stereo 3D Object Detection via Shape Prior Guided Instance Disparity Estimation. CVPR 2020.
116	SecAtten			45.76 %	55.24 %	42.75 %	0.1 s	1 core @ 2.5 Ghz (C/C++)

117	PCNet3D			45.66 %	56.08 %	42.95 %	0.05 s	GPU @ 3.5 Ghz (Python)

118	Disp R-CNN (velo)		code	45.66 %	63.16 %	41.14 %	0.387 s	GPU @ 2.5 Ghz (Python + C/C++)
J. Sun, L. Chen, Y. Xie, S. Zhang, Q. Jiang, X. Zhou and H. Bao: Disp R-CNN: Stereo 3D Object Detection via Shape Prior Guided Instance Disparity Estimation. CVPR 2020.
119	PillarHist			45.47 %	54.35 %	42.97 %	0.01 s	1 core @ 2.5 Ghz (C/C++)

120	HomoLoss(monoflex)		code	45.44 %	59.94 %	41.15 %	0.04 s	1 core @ 2.5 Ghz (Python)
J. Gu, B. Wu, L. Fan, J. Huang, S. Cao, Z. Xiang and X. Hua: Homography Loss for Monocular 3D Object Detection. Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) 2022.
121	GraphAlign(ICCV2023)		code	45.18 %	52.14 %	43.18 %	0.03 s	GPU @ 2.0 Ghz (Python)
Z. Song, H. Wei, L. Bai, L. Yang and C. Jia: GraphAlign: Enhancing accurate feature alignment by graph matching for multi-modal 3D object detection. Proceedings of the IEEE/CVF International Conference on Computer Vision 2023.
122	PCNet3D++			45.03 %	55.89 %	42.29 %	0.5 s	GPU @ 3.5 Ghz (Python)

123	Plane-Constraints		code	44.76 %	57.28 %	40.56 %	0.05 s	4 cores @ 3.0 Ghz (Python)
H. Yao, J. Chen, Z. Wang, X. Wang, X. Chai, Y. Qiu and P. Han: Vertex points are not enough: Monocular 3D object detection via intra-and inter-plane constraints. Neural Networks 2023.
124	focal			44.56 %	53.24 %	42.43 %	100 s	1 core @ 2.5 Ghz (Python)

125	MonoFlex			44.20 %	58.96 %	39.89 %	0.03 s	GPU @ 2.5 Ghz (Python)
Y. Zhang, J. Lu and J. Zhou: Objects are Different: Flexible Monocular 3D Object Detection. CVPR 2021.
126	AVOD-FPN		code	43.99 %	53.48 %	41.56 %	0.1 s	Titan X (Pascal)
J. Ku, M. Mozifian, J. Lee, A. Harakeh and S. Waslander: Joint 3D Proposal Generation and Object Detection from View Aggregation. IROS 2018.
127	PL++: PV-RCNN++			43.90 %	53.77 %	40.77 %	0.342 s	RTX 4060Ti (Python)

128	CAT-Det			43.86 %	52.75 %	41.15 %	0.3 s	GPU @ 2.5 Ghz (Python + C/C++)
Y. Zhang, J. Chen and D. Huang: CAT-Det: Contrastively Augmented Transformer for Multi-modal 3D Object Detection. CVPR 2022.
129	DSGN++		code	43.35 %	54.16 %	40.10 %	0.2 s	GeForce RTX 2080Ti
Y. Chen, S. Huang, S. Liu, B. Yu and J. Jia: DSGN++: Exploiting Visual-Spatial Relation for Stereo-Based 3D Detectors. IEEE Transactions on Pattern Analysis and Machine Intelligence 2022.
130	EPNet++			43.29 %	51.89 %	40.98 %	0.1 s	GPU @ 2.5 Ghz (Python)
Z. Liu, T. Huang, B. Li, X. Chen, X. Wang and X. Bai: EPNet++: Cascade Bi-Directional Fusion for Multi-Modal 3D Object Detection. IEEE Transactions on Pattern Analysis and Machine Intelligence 2022.
131	monospb			42.99 %	58.36 %	38.82 %	0.01 s	1 core @ 2.5 Ghz (Python)

132	Frustum-PointPillars		code	42.97 %	49.04 %	40.69 %	0.06 s	4 cores @ 3.0 Ghz (Python)
A. Paigwar, D. Sierra-Gonzalez, \. Erkent and C. Laugier: Frustum-PointPillars: A Multi-Stage Approach for 3D Object Detection using RGB Camera and LiDAR. International Conference on Computer Vision, ICCV, Workshop on Autonomous Vehicle Vision 2021.
133	MonoRCNN++		code	42.54 %	56.59 %	36.64 %	0.07 s	GPU @ 2.5 Ghz (Python)
X. Shi, Z. Chen and T. Kim: Multivariate Probabilistic Monocular 3D Object Detection. WACV 2023.
134	MonOAPC			42.52 %	56.84 %	38.43 %	0035 s	1 core @ 2.5 Ghz (Python)
H. Yao, J. Chen, Z. Wang, X. Wang, P. Han, X. Chai and Y. Qiu: Occlusion-Aware Plane-Constraints for Monocular 3D Object Detection. IEEE Transactions on Intelligent Transportation Systems 2023.
135	MonoPair			42.38 %	55.26 %	38.53 %	0.06 s	GPU @ 2.5 Ghz (Python + C/C++)
Y. Chen, L. Tai, K. Sun and M. Li: MonoPair: Monocular 3D Object Detection Using Pairwise Spatial Relationships. The IEEE Conference on Computer Vision and Pattern Recognition (CVPR) 2020.
136	T-SSD			41.55 %	52.44 %	39.16 %	0.04	1 core @ 2.0 Ghz (C/C++)

137	MonoDDE			41.09 %	55.28 %	36.85 %	0.04 s	1 core @ 2.5 Ghz (Python)
Z. Li, Z. Qu, Y. Zhou, J. Liu, H. Wang and L. Jiang: Diversity Matters: Fully Exploiting Depth Clues for Reliable Monocular 3D Object Detection. CVPR 2022.
138	mdab			41.04 %	56.10 %	36.74 %	0.02 s	1 core @ 2.5 Ghz (Python)

139	mdab			41.04 %	56.10 %	36.74 %	22 s	1 core @ 2.5 Ghz (C/C++)

140	PFF3D		code	40.99 %	48.75 %	38.99 %	0.05 s	GPU @ 3.0 Ghz (Python + C/C++)
L. Wen and K. Jo: Fast and Accurate 3D Object Detection for Lidar-Camera-Based Autonomous Vehicles Using One Shared Voxel-Based Backbone. IEEE Access 2021.
141	MMLAB LIGA-Stereo		code	40.98 %	53.16 %	38.12 %	0.4 s	1 core @ 2.5 Ghz (Python + C/C++)
X. Guo, S. Shi, X. Wang and H. Li: LIGA-Stereo: Learning LiDAR Geometry Aware Representations for Stereo-based 3D Detector. Proceedings of the IEEE/CVF International Conference on Computer Vision (ICCV) 2021.
142	DensePointPillars			40.62 %	48.97 %	38.14 %	0.03 s	GPU @ 2.5 Ghz (Python)

143	UniCuboid			40.07 %	52.95 %	36.09 %	0.1 s	1 core @ 2.5 Ghz (C/C++)

144	LPCG-Monoflex		code	39.79 %	56.60 %	35.42 %	0.03 s	1 core @ 2.5 Ghz (C/C++)
L. Peng, F. Liu, Z. Yu, S. Yan, D. Deng, Z. Yang, H. Liu and D. Cai: Lidar Point Cloud Guided Monocular 3D Object Detection. ECCV 2022.
145	AB3DMOT		code	39.76 %	50.30 %	36.90 %	0.0047s	1 core @ 2.5 Ghz (python)
X. Weng and K. Kitani: A Baseline for 3D Multi-Object Tracking. arXiv:1907.03961 2019.
146	SS3D			39.60 %	53.72 %	35.40 %	48 ms	Tesla V100 (Python)
E. Jörgensen, C. Zach and F. Kahl: Monocular 3D Object Detection and Box Fitting Trained End-to-End Using Intersection-over-Union Loss. CoRR 2019.
147	SemanticVoxels			38.95 %	45.59 %	37.21 %	0.04 s	GPU @ 2.5 Ghz (Python + C/C++)
J. Fei, W. Chen, P. Heidenreich, S. Wirges and C. Stiller: SemanticVoxels: Sequential Fusion for 3D Pedestrian Detection using LiDAR Point Cloud and Semantic Segmentation. MFI 2020.
148	MonoLiG		code	38.92 %	52.66 %	35.05 %	0.03 s	1 core @ 2.5 Ghz (C/C++)
A. Hekimoglu, M. Schmidt and A. Ramiro: Monocular 3D Object Detection with LiDAR Guided Semi Supervised Active Learning. 2023.
149	SeSame-voxel w/score		code	38.87 %	46.62 %	36.58 %	N/A s	GPU @ 1.5 Ghz (Python)
H. O, C. Yang and K. Huh: SeSame: Simple, Easy 3D Object Detection with Point-Wise Semantics. Proceedings of the Asian Conference on Computer Vision (ACCV) 2024.
150	geo-pillars			38.59 %	46.71 %	36.37 %	0.05 s	1 core @ 2.5 Ghz (C/C++)
ERROR: Wrong syntax in BIBTEX file.
151	DPM-VOC+VP			37.79 %	52.91 %	33.27 %	8 s	1 core @ 2.5 Ghz (C/C++)
B. Pepik, M. Stark, P. Gehler and B. Schiele: Multi-view and 3D Deformable Part Models. IEEE Transactions on Pattern Analysis and Machine Intelligence (TPAMI) 2015.
152	StereoDistill			37.58 %	48.49 %	34.41 %	0.4 s	1 core @ 2.5 Ghz (Python)
Z. Liu, X. Ye, X. Tan, D. Errui, Y. Zhou and X. Bai: StereoDistill: Pick the Cream from LiDAR for Distilling Stereo-based 3D Object Detection. Proceedings of the AAAI Conference on Artificial Intelligence 2023.
153	EQ-PVRCNN		code	36.49 %	43.67 %	34.67 %	0.2 s	GPU @ 2.5 Ghz (Python + C/C++)
Z. Yang, L. Jiang, Y. Sun, B. Schiele and J. Jia: A Unified Query-based Paradigm for Point Cloud Understanding. Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition 2022.
154	CG-Stereo			36.47 %	48.23 %	32.77 %	0.57 s	GeForce RTX 2080 Ti
C. Li, J. Ku and S. Waslander: Confidence Guided Stereo 3D Object Detection with Split Depth Estimation. IROS 2020.
155	TANet		code	36.21 %	42.54 %	34.39 %	0.035s	GPU @ 2.5 Ghz (Python + C/C++)
Z. Liu, X. Zhao, T. Huang, R. Hu, Y. Zhou and X. Bai: TANet: Robust 3D Object Detection from Point Clouds with Triple Attention. AAAI 2020.
156	YOLOStereo3D		code	35.62 %	48.99 %	31.58 %	0.1 s	GPU 1080Ti
Y. Liu, L. Wang and M. Liu: YOLOStereo3D: A Step Back to 2D for Efficient Stereo 3D Detection. 2021 International Conference on Robotics and Automation (ICRA) 2021.
157	SCNet			35.49 %	44.50 %	33.38 %	0.04 s	GPU @ 3.0 Ghz (Python)
Z. Wang, H. Fu, L. Wang, L. Xiao and B. Dai: SCNet: Subdivision Coding Network for Object Detection Based on 3D Point Cloud. IEEE Access 2019.
158	MonoDTR			35.11 %	49.41 %	31.41 %	0.04 s	1 core @ 2.5 Ghz (C/C++)
K. Huang, T. Wu, H. Su and W. Hsu: MonoDTR: Monocular 3D Object Detection with Depth-Aware Transformer. CVPR 2022.
159	BirdNet+		code	35.01 %	41.84 %	33.03 %	0.11 s	Titan Xp (PyTorch)
A. Barrera, J. Beltrán, C. Guindel, J. Iglesias and F. García: BirdNet+: Two-Stage 3D Object Detection in LiDAR through a Sparsity-Invariant Bird’s Eye View. IEEE Access 2021.
160	MonoEF			34.63 %	47.45 %	31.01 %	0.03 s	1 core @ 2.5 Ghz (Python)
Y. Zhou, Y. He, H. Zhu, C. Wang, H. Li and Q. Jiang: Monocular 3D Object Detection: An Extrinsic Parameter Free Approach. Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) 2021.
161	sensekitti		code	34.26 %	41.03 %	31.51 %	4.5 s	GPU @ 2.5 Ghz (Python + C/C++)
B. Yang, J. Yan, Z. Lei and S. Li: Craft Objects from Images. CVPR 2016.
162	D4LCN		code	33.62 %	46.73 %	28.71 %	0.2 s	GPU @ 2.5 Ghz (Python + C/C++)
M. Ding, Y. Huo, H. Yi, Z. Wang, J. Shi, Z. Lu and P. Luo: Learning Depth-Guided Convolutions for Monocular 3D Object Detection. CVPR 2020.
163	AEPF			33.46 %	39.98 %	31.35 %	0.05 s	GPU @ 2.5 Ghz (Python)

164	DDMP-3D			33.35 %	46.12 %	28.45 %	0.18 s	1 core @ 2.5 Ghz (Python)
L. Wang, L. Du, X. Ye, Y. Fu, G. Guo, X. Xue, J. Feng and L. Zhang: Depth-conditioned Dynamic Message Propagation for Monocular 3D Object Detection. CVPR 2020.
165	SparsePool		code	33.35 %	43.86 %	29.99 %	0.13 s	8 cores @ 2.5 Ghz (Python)
Z. Wang, W. Zhan and M. Tomizuka: Fusing bird view lidar point cloud and front view camera image for deep object detection. arXiv preprint arXiv:1711.06703 2017.
166	SparsePool		code	33.29 %	43.52 %	30.01 %	0.13 s	8 cores @ 2.5 Ghz (Python)
Z. Wang, W. Zhan and M. Tomizuka: Fusing bird view lidar point cloud and front view camera image for deep object detection. arXiv preprint arXiv:1711.06703 2017.
167	LSVM-MDPM-sv			33.01 %	45.60 %	29.27 %	10 s	4 cores @ 3.0 Ghz (C/C++)
P. Felzenszwalb, R. Girshick, D. McAllester and D. Ramanan: Object Detection with Discriminatively Trained Part-Based Models. PAMI 2010. A. Geiger, C. Wojek and R. Urtasun: Joint 3D Estimation of Objects and Scene Layout. NIPS 2011.
168	SeSame-pillar		code	32.73 %	40.30 %	30.56 %	N/A s	TITAN RTX @ 1.35 Ghz (Python)
H. O, C. Yang and K. Huh: SeSame: Simple, Easy 3D Object Detection with Point-Wise Semantics. Proceedings of the Asian Conference on Computer Vision (ACCV) 2024.
169	PointRGBNet			32.57 %	43.08 %	29.17 %	0.08 s	4 cores @ 2.5 Ghz (Python + C/C++)
P. Xie Desheng: Real-time Detection of 3D Objects Based on Multi-Sensor Information Fusion. Automotive Engineering 2022.
170	AVOD		code	32.19 %	42.54 %	29.09 %	0.08 s	Titan X (pascal)
J. Ku, M. Mozifian, J. Lee, A. Harakeh and S. Waslander: Joint 3D Proposal Generation and Object Detection from View Aggregation. IROS 2018.
171	Complexer-YOLO			32.13 %	37.32 %	28.94 %	0.06 s	GPU @ 3.5 Ghz (C/C++)
M. Simon, K. Amende, A. Kraus, J. Honer, T. Samann, H. Kaulbersch, S. Milz and H. Michael Gross: Complexer-YOLO: Real-Time 3D Object Detection and Tracking on Semantic Point Clouds. The IEEE Conference on Computer Vision and Pattern Recognition (CVPR) Workshops 2019.
172	RPN+BF		code	32.12 %	41.19 %	28.83 %	0.6 s	GPU @ 2.5 Ghz (Matlab + C/C++)
L. Zhang, L. Lin, X. Liang and K. He: Is Faster R-CNN Doing Well for Pedestrian Detection?. ECCV 2016.
173	DMF			32.00 %	39.86 %	30.12 %	0.2 s	1 core @ 2.5 Ghz (Python + C/C++)
X. J. Chen and W. Xu: Disparity-Based Multiscale Fusion Network for Transportation Detection. IEEE Transactions on Intelligent Transportation Systems 2022.
174	CMKD		code	31.97 %	42.60 %	29.13 %	0.1 s	1 core @ 2.5 Ghz (C/C++)
Y. Hong, H. Dai and Y. Ding: Cross-Modality Knowledge Distillation Network for Monocular 3D Object Detection. ECCV 2022.
175	M3D-RPN		code	31.88 %	44.33 %	28.55 %	0.16 s	GPU @ 1.5 Ghz (Python)
G. Brazil and X. Liu: M3D-RPN: Monocular 3D Region Proposal Network for Object Detection . ICCV 2019 .
176	Point-GNN		code	31.86 %	39.16 %	29.65 %	0.6 s	GPU @ 2.5 Ghz (Python)
W. Shi and R. Rajkumar: Point-GNN: Graph Neural Network for 3D Object Detection in a Point Cloud. CVPR 2020.
177	SubCat			31.26 %	42.31 %	27.39 %	1.2 s	6 cores @ 2.5 Ghz (Matlab + C/C++)
E. Ohn-Bar and M. Trivedi: Fast and Robust Object Detection Using Visual Subcategories. Computer Vision and Pattern Recognition Workshops Mobile Vision 2014.
178	SeSame-pillar w/scor		code	30.83 %	38.16 %	28.98 %	N/A s	1 core @ 2.5 Ghz (C/C++)
H. O, C. Yang and K. Huh: SeSame: Simple, Easy 3D Object Detection with Point-Wise Semantics. Proceedings of the Asian Conference on Computer Vision (ACCV) 2024.
179	Fade-kd			30.50 %	38.03 %	28.83 %	0.01 s	1 core @ 2.5 Ghz (C/C++)

180	Aug3D-RPN			29.75 %	40.50 %	25.96 %	0.08 s	1 core @ 2.5 Ghz (C/C++)
C. He, J. Huang, X. Hua and L. Zhang: Aug3D-RPN: Improving Monocular 3D Object Detection by Synthetic Images with Virtual Depth. 2021.
181	BirdNet+ (legacy)		code	29.56 %	36.76 %	28.10 %	0.1 s	Titan Xp (PyTorch)
A. Barrera, C. Guindel, J. Beltrán and F. García: BirdNet+: End-to-End 3D Object Detection in LiDAR Bird’s Eye View. 2020 IEEE 23rd International Conference on Intelligent Transportation Systems (ITSC) 2020.
182	SeSame-point w/score		code	28.86 %	39.33 %	26.47 %	N/A s	1 core @ 1.5 Ghz (Python)
H. O, C. Yang and K. Huh: SeSame: Simple, Easy 3D Object Detection with Point-Wise Semantics. Proceedings of the Asian Conference on Computer Vision (ACCV) 2024.
183	RT3D-GMP			28.75 %	40.81 %	25.13 %	0.06 s	GPU @ 2.5 Ghz (Python + C/C++)
H. Königshof and C. Stiller: Learning-Based Shape Estimation with Grid Map Patches for Realtime 3D Object Detection for Automated Driving. 2020 IEEE 23rd International Conference on Intelligent Transportation Systems (ITSC) 2020.
184	CMAN			28.16 %	40.27 %	24.82 %	0.15 s	1 core @ 2.5 Ghz (Python)
C. Yuanzhouhan Cao: CMAN: Leaning Global Structure Correlation for Monocular 3D Object Detection. IEEE Trans. Intell. Transport. Syst. 2022.
185	Cube R-CNN		code	28.07 %	34.26 %	25.14 %	0.05 s	GPU @ 2.5 Ghz (Python)
G. Brazil, A. Kumar, J. Straub, N. Ravi, J. Johnson and G. Gkioxari: Omni3D: A Large Benchmark and Model for 3D Object Detection in the Wild. CVPR 2023.
186	CIE			27.84 %	37.65 %	25.24 %	0.1 s	1 core @ 2.5 Ghz (C/C++)
Anonymities: Consistency of Implicit and Explicit Features Matters for Monocular 3D Object Detection. arXiv preprint arXiv:2207.07933 2022.
187	PGD-FCOS3D		code	27.61 %	40.20 %	24.29 %	0.03 s	1 core @ 2.5 Ghz (C/C++)
T. Wang, X. Zhu, J. Pang and D. Lin: Probabilistic and Geometric Depth: Detecting Objects in Perspective. Conference on Robot Learning (CoRL) 2021.
188	FMF-occlusion-net			26.28 %	38.13 %	22.91 %	0.16 s	1 core @ 2.5 Ghz (Python + C/C++)
H. Liu, H. Liu, Y. Wang, F. Sun and W. Huang: Fine-grained Multi-level Fusion for Anti- occlusion Monocular 3D Object Detection. IEEE Transactions on Image Processing 2022.
189	AMNet+DDAD15M		code	25.45 %	33.83 %	22.95 %	0.03 s	1 core @ 2.5 Ghz (C/C++)
H. Pan, Y. Jia, J. Wang and W. Sun: MonoAMNet: Three-Stage Real-Time Monocular 3D Object Detection With Adaptive Methods. IEEE Transactions on Intelligent Transportation Systems 2025.
190	MonoFRD			24.92 %	33.47 %	22.38 %	0.01 s	1 core @ 2.5 Ghz (C/C++)
Z. Gong, Y. Zhao, F. Zhang, G. Gui, B. Chen, L. Yu, H. Wang, C. Yang and W. Gui: Color intuitive feature guided depth-height fusion and volume rendering for monocular 3D object detection. IEEE Transactions on Intelligent Vehicles(Major Revison) 2024.
191	DFR-Net			24.88 %	35.75 %	21.72 %	0.18 s	1080 Ti (Pytorch)
Z. Zou, X. Ye, L. Du, X. Cheng, X. Tan, L. Zhang, J. Feng, X. Xue and E. Ding: The devil is in the task: Exploiting reciprocal appearance-localization features for monocular 3d object detection . ICCV 2021.
192	DSGN		code	24.32 %	31.21 %	23.09 %	0.67 s	NVIDIA Tesla V100
Y. Chen, S. Liu, X. Shen and J. Jia: DSGN: Deep Stereo Geometry Network for 3D Object Detection. CVPR 2020.
193	ACF			24.31 %	32.23 %	21.70 %	1 s	1 core @ 3.5 Ghz (Matlab + C/C++)
P. Doll\'ar, R. Appel, S. Belongie and P. Perona: Fast Feature Pyramids for Object Detection. PAMI 2014.
194	PS-SVDM			24.19 %	33.74 %	21.63 %	1 s	1 core @ 2.5 Ghz (Python)
Y. Shi: SVDM: Single-View Diffusion Model for Pseudo-Stereo 3D Object Detection. arXiv preprint arXiv:2307.02270 2023.
195	PS-fld		code	23.67 %	32.84 %	21.40 %	0.25 s	1 core @ 2.5 Ghz (C/C++)
Y. Chen, H. Dai and Y. Ding: Pseudo-Stereo for Monocular 3D Object Detection in Autonomous Driving. Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) 2022.
196	SGM3D		code	23.54 %	33.73 %	20.50 %	0.03 s	1 core @ 2.5 Ghz (C/C++)
Z. Zhou, L. Du, X. Ye, Z. Zou, X. Tan, L. Zhang, X. Xue and J. Feng: SGM3D: Stereo Guided Monocular 3D Object Detection. RA-L 2022.
197	AMNet		code	23.48 %	31.60 %	21.10 %	0.03 s	GPU @ 1.0 Ghz (Python)
H. Pan, Y. Jia, J. Wang and W. Sun: MonoAMNet: Three-Stage Real-Time Monocular 3D Object Detection With Adaptive Methods. IEEE Transactions on Intelligent Transportation Systems 2025.
198	PS-SVDM			23.28 %	32.63 %	20.76 %	1 s	1 core @ 2.5 Ghz (Python)
Y. Shi: SVDM: Single-View Diffusion Model for Pseudo-Stereo 3D Object Detection. arXiv preprint arXiv:2307.02270 2023.
199	Fade			23.24 %	30.04 %	21.82 %	0.01 s	1 core @ 2.5 Ghz (C/C++)

200	multi-task CNN			22.80 %	30.30 %	20.47 %	25.1 ms	GPU @ 2.0 Ghz (Python)
M. Oeljeklaus, F. Hoffmann and T. Bertram: A Fast Multi-Task CNN for Spatial Understanding of Traffic Scenes. IEEE Intelligent Transportation Systems Conference 2018.
201	ACF-MR			22.61 %	29.23 %	20.08 %	0.6 s	1 core @ 3.5 Ghz (C/C++)
R. Rajaram, E. Ohn-Bar and M. Trivedi: Looking at Pedestrians at Different Scales: A Multi-resolution Approach and Evaluations. T-ITS 2016.
202	OC Stereo		code	22.02 %	31.36 %	20.20 %	0.35 s	1 core @ 2.5 Ghz (Python + C/C++)
A. Pon, J. Ku, C. Li and S. Waslander: Object-Centric Stereo Matching for 3D Object Detection. ICRA 2020.
203	BirdNet			21.83 %	27.12 %	20.56 %	0.11 s	Titan Xp (Caffe)
J. Beltrán, C. Guindel, F. Moreno, D. Cruzado, F. García and A. Escalera: BirdNet: A 3D Object Detection Framework from LiDAR Information. 2018 21st International Conference on Intelligent Transportation Systems (ITSC) 2018.
204	MonoNeRD		code	20.54 %	28.43 %	18.36 %	na s	1 core @ 2.5 Ghz (Python)
J. Xu, L. Peng, H. Cheng, H. Li, W. Qian, K. Li, W. Wang and D. Cai: MonoNeRD: NeRF-like Representations for Monocular 3D Object Detection. ICCV 2023.
205	DPM-C8B1			19.17 %	27.79 %	16.48 %	15 s	4 cores @ 2.5 Ghz (Matlab + C/C++)
J. Yebes, L. Bergasa and M. García-Garrido: Visual Object Recognition with 3D-Aware Features in KITTI Urban Scenes. Sensors 2015. J. Yebes, L. Bergasa, R. Arroyo and A. Lázaro: Supervised learning and evaluation of KITTI's cars detector with DPM. IV 2014.
206	ESGN			19.17 %	26.02 %	16.90 %	0.06 s	GPU @ 2.5 Ghz (Python + C/C++)
A. Gao, Y. Pang, J. Nie, Z. Shao, J. Cao, Y. Guo and X. Li: ESGN: Efficient Stereo Geometry Network for Fast 3D Object Detection. IEEE Transactions on Circuits and Systems for Video Technology 2022.
207	LLW			18.96 %	26.40 %	16.95 %	0.1 s	1 core @ 2.5 Ghz (C/C++)

208	RefinedMPL			17.26 %	25.83 %	15.41 %	0.15 s	GPU @ 2.5 Ghz (Python + C/C++)
J. Vianney, S. Aich and B. Liu: RefinedMPL: Refined Monocular PseudoLiDAR for 3D Object Detection in Autonomous Driving. arXiv preprint arXiv:1911.09712 2019.
209	CaDDN		code	17.13 %	24.45 %	15.79 %	0.63 s	GPU @ 2.5 Ghz (Python)
C. Reading, A. Harakeh, J. Chae and S. Waslander: Categorical Depth Distribution Network for Monocular 3D Object Detection. CVPR 2021.
210	RT3DStereo			15.34 %	21.41 %	13.23 %	0.08 s	GPU @ 2.5 Ghz (C/C++)
H. Königshof, N. Salscheider and C. Stiller: Realtime 3D Object Detection for Automated Driving Using Stereo Vision and Semantic Information. Proc. IEEE Intl. Conf. Intelligent Transportation Systems 2019.
211	CPD++(unsupervised)		code	10.69 %	12.16 %	10.20 %	0.1 s	GPU @ >3.5 Ghz (Python)

212	CPD(unsupervised)		code	5.31 %	6.68 %	4.89 %	0.1 s	GPU @ >3.5 Ghz (Python + C/C++)
H. Wu, S. Zhao, X. Huang, C. Wen, X. Li and C. Wang: Commonsense Prototype for Outdoor Unsupervised 3D Object Detection. CVPR 2024.
213	GATE3D		code	0.80 %	1.15 %	0.86 %	0.01 s	1 core @ 2.5 Ghz (C/C++)

214	f3sd		code	0.01 %	0.01 %	0.01 %	1.67 s	1 core @ 2.5 Ghz (C/C++)

215	DA3D+KM3D+v2-99		code	0.00 %	0.00 %	0.00 %	0.120s	GPU @ 2.5 Ghz (Python)
Y. Jia, J. Wang, H. Pan and W. Sun: Enhancing Monocular 3-D Object Detection Through Data Augmentation Strategies. IEEE Transactions on Instrumentation and Measurement 2024.
216	DA3D+KM3D		code	0.00 %	0.00 %	0.00 %	0.02 s	GPU @ 2.5 Ghz (Python)
Y. Jia, J. Wang, H. Pan and W. Sun: Enhancing Monocular 3-D Object Detection Through Data Augmentation Strategies. IEEE Transactions on Instrumentation and Measurement 2024.
217	DA3D		code	0.00 %	0.00 %	0.00 %	0.03 s	1 core @ 2.5 Ghz (Python)
Y. Jia, J. Wang, H. Pan and W. Sun: Enhancing Monocular 3-D Object Detection Through Data Augmentation Strategies. IEEE Transactions on Instrumentation and Measurement 2024.

	Method	Setting	Code	Moderate	Easy	Hard	Runtime	Environment
1	LCANet			85.95 %	91.94 %	78.58 %	1 s	1 core @ 2.5 Ghz (C/C++)

2	TED		code	84.08 %	92.46 %	78.07 %	0.1 s	1 core @ 2.5 Ghz (C/C++)
H. Wu, C. Wen, W. Li, R. Yang and C. Wang: Transformation-Equivariant 3D Object Detection for Autonomous Driving. AAAI 2023.
3	CasA++		code	83.98 %	92.24 %	78.05 %	0.1 s	1 core @ 2.5 Ghz (C/C++)
H. Wu, J. Deng, C. Wen, X. Li and C. Wang: CasA: A Cascade Attention Network for 3D Object Detection from LiDAR point clouds. IEEE Transactions on Geoscience and Remote Sensing 2022.
4	UPIDet		code	83.78 %	89.86 %	76.98 %	0.11 s	1 core @ 2.5 Ghz (C/C++)
Y. Zhang, Q. Zhang, J. Hou, Y. Yuan and G. Xing: Unleash the Potential of Image Branch for Cross-modal 3D Object Detection. Thirty-seventh Conference on Neural Information Processing Systems 2023.
5	LoGoNet		code	83.51 %	89.90 %	77.41 %	0.1 s	1 core @ 2.5 Ghz (C/C++)
X. Li, T. Ma, Y. Hou, B. Shi, Y. Yang, Y. Liu, X. Wu, Q. Chen, Y. Li, Y. Qiao and others: LoGoNet: Towards Accurate 3D Object Detection with Local-to-Global Cross-Modal Fusion. CVPR 2023.
6	CasA		code	82.95 %	92.71 %	76.78 %	0.1 s	1 core @ 2.5 Ghz (C/C++)
H. Wu, J. Deng, C. Wen, X. Li and C. Wang: CasA: A Cascade Attention Network for 3D Object Detection from LiDAR point clouds. IEEE Transactions on Geoscience and Remote Sensing 2022.
7	RangeIoUDet			81.24 %	90.24 %	74.49 %	0.02 s	GPU @ 2.5 Ghz (Python + C/C++)
Z. Liang, Z. Zhang, M. Zhang, X. Zhao and S. Pu: RangeIoUDet: Range Image Based Real-Time 3D Object Detector Optimized by Intersection Over Union. CVPR 2021.
8	MPC3DNet			81.24 %	88.04 %	74.58 %	0.05 s	GPU @ 1.5 Ghz (Python)

9	HMFI		code	81.13 %	89.09 %	74.30 %	0.1 s	1 core @ 2.5 Ghz (C/C++)
X. Li, B. Shi, Y. Hou, X. Wu, T. Ma, Y. Li and L. He: Homogeneous Multi-modal Feature Fusion and Interaction for 3D Object Detection. ECCV 2022.
10	MLF-DET			81.07 %	87.17 %	73.92 %	0.09 s	1 core @ 2.5 Ghz (C/C++)
Z. Lin, Y. Shen, S. Zhou, S. Chen and N. Zheng: MLF-DET: Multi-Level Fusion for Cross- Modal 3D Object Detection. International Conference on Artificial Neural Networks 2023.
11	USVLab BSAODet		code	80.87 %	86.64 %	73.87 %	0.04 s	1 core @ 2.5 Ghz (C/C++)
W. Xiao, Y. Peng, C. Liu, J. Gao, Y. Wu and X. Li: Balanced Sample Assignment and Objective for Single-Model Multi-Class 3D Object Detection. IEEE Transactions on Circuits and Systems for Video Technology 2023.
12	AFFN			80.36 %	89.76 %	74.40 %	0.5 s	1 core @ 2.5 Ghz (C/C++)

13	GEFPN			80.36 %	89.76 %	74.40 %	0.5 s	1 core @ 2.5 Ghz (C/C++)

14	CAT-Det			80.25 %	87.79 %	73.41 %	0.3 s	GPU @ 2.5 Ghz (Python + C/C++)
Y. Zhang, J. Chen and D. Huang: CAT-Det: Contrastively Augmented Transformer for Multi-modal 3D Object Detection. CVPR 2022.
15	RagNet3D		code	80.23 %	89.25 %	73.69 %	0.05 s	1 core @ 2.5 Ghz (C/C++)
J. Chen, Y. Han, Z. Yan, J. Qian, J. Li and J. Yang: Ragnet3d: Learning Distinguishable Representation for Pooled Grids in 3d Object Detection. Available at SSRN 4979473 .
16	EQ-PVRCNN		code	80.09 %	88.92 %	73.79 %	0.2 s	GPU @ 2.5 Ghz (Python + C/C++)
Z. Yang, L. Jiang, Y. Sun, B. Schiele and J. Jia: A Unified Query-based Paradigm for Point Cloud Understanding. Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition 2022.
17	DSA-PV-RCNN		code	80.05 %	88.52 %	74.20 %	0.08 s	1 core @ 2.5 Ghz (Python + C/C++)
P. Bhattacharyya, C. Huang and K. Czarnecki: SA-Det3D: Self-Attention Based Context-Aware 3D Object Detection. 2021.
18	HINTED		code	79.73 %	86.59 %	73.13 %	0.04 s	1 core @ 2.5 Ghz (C/C++)
Q. Xia, W. Ye, H. Wu, S. Zhao, L. Xing, X. Huang, J. Deng, X. Li, C. Wen and C. Wang: HINTED: Hard Instance Enhanced Detector with Mixed-Density Feature Fusion for Sparsely- Supervised 3D Object Detection. Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition 2024.
19	LPFusion_three_class			79.70 %	87.29 %	70.82 %	0.1 s	1 core @ 2.5 Ghz (C/C++)

20	MMLab PV-RCNN		code	79.70 %	86.43 %	72.96 %	0.08 s	1 core @ 2.5 Ghz (Python + C/C++)
S. Shi, C. Guo, L. Jiang, Z. Wang, J. Shi, X. Wang and H. Li: PV-RCNN: Point-Voxel Feature Set Abstraction for 3D Object Detection. CVPR 2020.
21	test			79.67 %	91.10 %	73.39 %	0.04 s	GPU @ 1.5 Ghz (Python + C/C++)

22	PDV		code	79.34 %	88.66 %	72.56 %	0.1 s	1 core @ 2.5 Ghz (C/C++)
J. Hu, T. Kuai and S. Waslander: Point Density-Aware Voxels for LiDAR 3D Object Detection. CVPR 2022.
23	vsis-PHNet			79.33 %	90.53 %	74.11 %	0.03 s	1 core @ 2.5 Ghz (C/C++)

24	SCNet3D			79.23 %	89.47 %	72.30 %	0.08 s	1 core @ 2.5 Ghz (Python)

25	PASS-PV-RCNN-Plus			78.82 %	86.15 %	72.28 %	1 s	1 core @ 2.5 Ghz (Python)
Anonymous: Leveraging Anchor-based LiDAR 3D Object Detection via Point Assisted Sample Selection. will submit to computer vision conference/journal 2024.
26	M3DeTR		code	78.80 %	87.21 %	71.88 %	n/a s	GPU @ 1.0 Ghz (Python)
T. Guan, J. Wang, S. Lan, R. Chandra, Z. Wu, L. Davis and D. Manocha: M3DeTR: Multi-representation, Multi- scale, Mutual-relation 3D Object Detection with Transformers. 2021.
27	BVIFusion+			78.77 %	88.03 %	72.02 %	0.09 s	1 core @ 2.5 Ghz (Python)

28	IA-SSD (single)		code	78.34 %	88.78 %	71.63 %	0.013 s	1 core @ 2.5 Ghz (C/C++)
Y. Zhang, Q. Hu, G. Xu, Y. Ma, J. Wan and Y. Guo: Not All Points Are Equal: Learning Highly Efficient Point-based Detectors for 3D LiDAR Point Clouds. CVPR 2022.
29	PHNetc			78.34 %	89.21 %	73.27 %	0.03 s	1 core @ 2.5 Ghz (C/C++)

30	HotSpotNet			78.31 %	85.79 %	71.24 %	0.04 s	1 core @ 2.5 Ghz (Python + C/C++)
Q. Chen, L. Sun, Z. Wang, K. Jia and A. Yuille: object as hotspots. Proceedings of the European Conference on Computer Vision (ECCV) 2020.
31	PG-RCNN		code	78.30 %	87.89 %	71.76 %	0.06 s	GPU @ 1.5 Ghz (Python)
I. Koo, I. Lee, S. Kim, H. Kim, W. Jeon and C. Kim: PG-RCNN: Semantic Surface Point Generation for 3D Object Detection. 2023.
32	AFFN-G			77.75 %	84.80 %	72.57 %	0.5 s	1 core @ 2.5 Ghz (C/C++)

33	dsvd+vx			77.68 %	91.83 %	70.46 %	0.1 s	1 core @ 2.5 Ghz (C/C++)

34	FIRM-Net_SCF+			77.53 %	89.53 %	70.52 %	0.07 s	1 core @ 2.5 Ghz (Python)

35	MMLab-PartA^2		code	77.52 %	88.70 %	70.41 %	0.08 s	GPU @ 2.5 Ghz (Python + C/C++)
S. Shi, Z. Wang, J. Shi, X. Wang and H. Li: From Points to Parts: 3D Object Detection from Point Cloud with Part-aware and Part-aggregation Network. IEEE Transactions on Pattern Analysis and Machine Intelligence 2020.
36	MFB3D			77.45 %	86.12 %	71.45 %	0.14 s	1 core @ 2.5 Ghz (Python)

37	DFAF3D			77.41 %	86.98 %	70.42 %	0.05 s	1 core @ 2.5 Ghz (Python)
Q. Tang, X. Bai, J. Guo, B. Pan and W. Jiang: DFAF3D: A dual-feature-aware anchor-free single-stage 3D detector for point clouds. Image and Vision Computing 2023.
38	FIRM-Net-SCF			77.29 %	89.45 %	70.33 %	0.07 s	1 core @ 2.5 Ghz (Python)

39	PointPainting			76.92 %	87.33 %	68.21 %	0.4 s	GPU @ 2.5 Ghz (Python + C/C++)
S. Vora, A. Lang, B. Helou and O. Beijbom: PointPainting: Sequential Fusion for 3D Object Detection. CVPR 2020.
40	3ONet			76.91 %	88.98 %	69.85 %	0.1 s	1 core @ 2.5 Ghz (C/C++)
H. Hoang and M. Yoo: 3ONet: 3-D Detector for Occluded Object Under Obstructed Conditions. IEEE Sensors Journal 2023.
41	OGMMDet		code	76.86 %	86.40 %	71.57 %	0.01 s	1 core @ 2.5 Ghz (C/C++)

42	ANM		code	76.86 %	88.28 %	71.57 %	ANM	ANM

43	GraphAlign(ICCV2023)		code	76.81 %	84.53 %	71.90 %	0.03 s	GPU @ 2.0 Ghz (Python)
Z. Song, H. Wei, L. Bai, L. Yang and C. Jia: GraphAlign: Enhancing accurate feature alignment by graph matching for multi-modal 3D object detection. Proceedings of the IEEE/CVF International Conference on Computer Vision 2023.
44	AFFN-Ga			76.77 %	84.70 %	70.83 %	0.5 s	1 core @ 2.5 Ghz (C/C++)

45	F-ConvNet		code	76.71 %	86.39 %	66.92 %	0.47 s	GPU @ 2.5 Ghz (Python + C/C++)
Z. Wang and K. Jia: Frustum ConvNet: Sliding Frustums to Aggregate Local Point-Wise Features for Amodal 3D Object Detection. IROS 2019.
46	P2V-RCNN			76.52 %	88.21 %	69.90 %	0.1 s	2 cores @ 2.5 Ghz (Python)
J. Li, S. Luo, Z. Zhu, H. Dai, A. Krylov, Y. Ding and L. Shao: P2V-RCNN: Point to Voxel Feature Learning for 3D Object Detection from Point Clouds. IEEE Access 2021.
47	CGML			76.40 %	85.45 %	70.98 %	0.33 s	1 core @ 2.5 Ghz (C/C++)

48	SA V1			76.01 %	86.88 %	69.23 %	0.5 s	GPU @ 2.5 Ghz (Python)

49	CG-SSD			75.79 %	85.23 %	69.60 %	0.01 s	1 core @ 2.5 Ghz (C/C++)

50	focal			75.67 %	83.97 %	69.88 %	100 s	1 core @ 2.5 Ghz (Python)

51	GeVo			75.67 %	83.97 %	69.88 %	0.05 s	1 core @ 2.5 Ghz (C/C++)

52	CAIA_PRO		code	75.64 %	86.57 %	69.16 %	0.01 s	1 core @ 2.5 Ghz (C/C++)

53	ACFNet			75.34 %	86.11 %	70.41 %	0.11 s	1 core @ 2.5 Ghz (C/C++)
Y. Tian, X. Zhang, X. Wang, J. Xu, J. Wang, R. Ai, W. Gu and W. Ding: ACF-Net: Asymmetric Cascade Fusion for 3D Detection With LiDAR Point Clouds and Images. IEEE Transactions on Intelligent Vehicles 2023.
54	MLFusion-VS			74.99 %	83.55 %	69.20 %	0.06 s	1 core @ 2.5 Ghz (C/C++)

55	Fast-CLOCs			74.74 %	89.54 %	67.54 %	0.1 s	GPU @ 2.5 Ghz (Python)
S. Pang, D. Morris and H. Radha: Fast-CLOCs: Fast Camera-LiDAR Object Candidates Fusion for 3D Object Detection. Proceedings of the IEEE/CVF Winter Conference on Applications of Computer Vision (WACV) 2022.
56	GSG-FPS		code	74.72 %	87.44 %	68.28 %	0.01 s	1 core @ 2.5 Ghz (C/C++)

57	SVGA-Net			74.64 %	84.62 %	67.64 %	0.03s	1 core @ 2.5 Ghz (Python + C/C++)
Q. He, Z. Wang, H. Zeng, Y. Zeng and Y. Liu: SVGA-Net: Sparse Voxel-Graph Attention Network for 3D Object Detection from Point Clouds. AAAI 2022.
58	Res3DNet			74.52 %	88.38 %	67.89 %	0.05 s	GPU @ 3.5 Ghz (Python)

59	ACDet		code	74.52 %	88.21 %	68.33 %	0.05 s	1 core @ 2.5 Ghz (C/C++)
J. Xu, G. Wang, X. Zhang and G. Wan: ACDet: Attentive Cross-view Fusion for LiDAR-based 3D Object Detection. 3DV 2022.
60	SVFMamba		code	74.33 %	83.36 %	68.05 %	N/A s	1 core @ 2.5 Ghz (C/C++)

61	Voxel RCNN-Focal*		code	74.14 %	81.58 %	68.93 %	0.2 s	1 core @ 2.5 Ghz (Python)

62	LumiNet		code	74.03 %	87.99 %	67.13 %	0.1 s	1 core @ 2.5 Ghz (Python)

63	MSMA V1			73.69 %	85.31 %	67.38 %	0.5 s	GPU @ 2.5 Ghz (Python)

64	VPFNet		code	73.62 %	82.08 %	65.27 %	0.2 s	1 core @ 2.5 Ghz (C/C++)
C. Wang, H. Chen and L. Fu: VPFNet: Voxel-Pixel Fusion Network for Multi-class 3D Object Detection. 2021. C. Wang, H. Chen, Y. Chen, P. Hsiao and L. Fu: VoPiFNet: Voxel-Pixel Fusion Network for Multi-Class 3D Object Detection. IEEE Transactions on Intelligent Transportation Systems 2024.
65	SFA_IGCL_Focalsconv*		code	73.55 %	80.36 %	67.90 %	0.2 s	1 core @ 2.5 Ghz (C/C++)

66	DVFENet			73.43 %	85.32 %	66.87 %	0.05 s	1 core @ 2.5 Ghz (Python + C/C++)
Y. He, G. Xia, Y. Luo, L. Su, Z. Zhang, W. Li and P. Wang: DVFENet: Dual-branch Voxel Feature Extraction Network for 3D Object Detection. Neurocomputing 2021.
67	FocalsConv*			73.37 %	83.11 %	68.60 %	0.2 s	1 core @ 2.5 Ghz (C/C++)

68	SRDL			73.21 %	85.22 %	66.45 %	0.05 s	1 core @ 2.5 Ghz (Python + C/C++)
ERROR: Wrong syntax in BIBTEX file.
69	L-AUG			73.07 %	83.69 %	67.72 %	0.1 s	1 core @ 2.5 Ghz (C/C++)
T. Cortinhal, I. Gouigah and E. Aksoy: Semantics-aware LiDAR-Only Pseudo Point Cloud Generation for 3D Object Detection. 2023.
70	MMLab-PointRCNN		code	72.81 %	85.94 %	65.84 %	0.1 s	GPU @ 2.5 Ghz (Python + C/C++)
S. Shi, X. Wang and H. Li: Pointrcnn: 3d object proposal generation and detection from point cloud. Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition 2019.
71	SIF			72.73 %	84.96 %	64.94 %	0.1 s	1 core @ 2.5 Ghz (C/C++)
P. An: SIF. Submitted to CVIU 2021.
72	XView			72.70 %	87.59 %	64.96 %	0.1 s	1 core @ 2.5 Ghz (C/C++)
L. Xie, G. Xu, D. Cai and X. He: X-view: Non-egocentric Multi-View 3D Object Detector. 2021.
73	FromVoxelToPoint		code	72.62 %	86.71 %	65.42 %	0.1 s	1 core @ 2.5 Ghz (C/C++)
J. Li, H. Dai, L. Shao and Y. Ding: From Voxel to Point: IoU-guided 3D Object Detection for Point Cloud with Voxel-to- Point Decoder. MM '21: The 29th ACM International Conference on Multimedia (ACM MM) 2021.
74	R50_SACINet			72.43 %	84.30 %	66.20 %	0.06 s	1 core @ 2.5 Ghz (C/C++)

75	EOTL		code	72.37 %	82.07 %	62.06 %	TBD s	1 core @ 2.5 Ghz (Python + C/C++)
R. Yang, Z. Yan, T. Yang, Y. Wang and Y. Ruichek: Efficient Online Transfer Learning for Road Participants Detection in Autonomous Driving. IEEE Sensors Journal 2023.
76	H^23D R-CNN		code	72.20 %	85.09 %	65.25 %	0.03 s	1 core @ 2.5 Ghz (C/C++)
J. Deng, W. Zhou, Y. Zhang and H. Li: From Multi-View to Hollow-3D: Hallucinated Hollow-3D R-CNN for 3D Object Detection. IEEE Transactions on Circuits and Systems for Video Technology 2021.
77	Test_dif		code	72.16 %	85.41 %	67.13 %	0.01 s	1 core @ 2.5 Ghz (C/C++)

78	SFA-GCL(80)		code	71.96 %	86.93 %	64.97 %	0.04 s	1 core @ 2.5 Ghz (C/C++)

79	HA-PillarNet			71.91 %	81.85 %	67.07 %	0.05 s	1 core @ 2.5 Ghz (Python)

80	SeSame-point		code	71.88 %	83.97 %	65.00 %	N/A s	TITAN RTX @ 1.35 Ghz (Python)
H. O, C. Yang and K. Huh: SeSame: Simple, Easy 3D Object Detection with Point-Wise Semantics. Proceedings of the Asian Conference on Computer Vision (ACCV) 2024.
81	L_SACINet			71.87 %	81.49 %	65.56 %	0.04 s	1 core @ 2.5 Ghz (C/C++)

82	SFA-GCL(80, k=4)		code	71.75 %	86.48 %	62.80 %	0.04 s	1 core @ 2.5 Ghz (C/C++)

83	SFA-GCL		code	71.73 %	86.78 %	64.73 %	0.04 s	1 core @ 2.5 Ghz (C/C++)

84	HMNet			71.60 %	85.16 %	65.66 %	0.1 s	1 core @ 2.5 Ghz (C/C++)

85	S-AT GCN			71.04 %	82.31 %	65.13 %	0.02 s	GPU @ 2.0 Ghz (Python)
L. Wang, C. Wang, X. Zhang, T. Lan and J. Li: S-AT GCN: Spatial-Attention Graph Convolution Network based Feature Enhancement for 3D Object Detection. CoRR 2021.
86	Anonymous			70.64 %	82.45 %	65.34 %	0.02 s	GPU @ 2.5 Ghz (Python)

87	R2Pfusion-Det			70.51 %	82.65 %	63.37 %	0.3 s	1 core @ 2.5 Ghz (C/C++)

88	SecAtten			70.33 %	80.62 %	64.31 %	0.1 s	1 core @ 2.5 Ghz (C/C++)

89	MGAF-3DSSD		code	70.16 %	86.28 %	62.99 %	0.1 s	1 core @ 2.5 Ghz (C/C++)
J. Li, H. Dai, L. Shao and Y. Ding: Anchor-free 3D Single Stage Detector with Mask-Guided Attention for Point Cloud. MM '21: The 29th ACM International Conference on Multimedia (ACM MM) 2021.
90	IA-SSD (multi)		code	70.13 %	84.82 %	65.13 %	0.014 s	1 core @ 2.5 Ghz (C/C++)
Y. Zhang, Q. Hu, G. Xu, Y. Ma, J. Wan and Y. Guo: Not All Points Are Equal: Learning Highly Efficient Point-based Detectors for 3D LiDAR Point Clouds. CVPR 2022.
91	PCNet3D++			69.59 %	82.16 %	63.42 %	0.5 s	GPU @ 3.5 Ghz (Python)

92	AB3DMOT		code	69.54 %	82.18 %	62.98 %	0.0047s	1 core @ 2.5 Ghz (Python)
X. Weng and K. Kitani: A Baseline for 3D Multi-Object Tracking. arXiv:1907.03961 2019.
93	SFA-GCL(baseline)		code	69.29 %	86.10 %	60.34 %	0.04 s	1 core @ 2.5 Ghz (C/C++)

94	SeSame-voxel		code	69.21 %	86.97 %	62.47 %	N/A s	TITAN RTX @ 1.35 Ghz (Python)
H. O, C. Yang and K. Huh: SeSame: Simple, Easy 3D Object Detection with Point-Wise Semantics. Proceedings of the Asian Conference on Computer Vision (ACCV) 2024.
95	SFA-GCL_dataaug		code	69.10 %	85.79 %	62.14 %	0.04 s	1 core @ 2.5 Ghz (C/C++)

96	SFA-GCL		code	69.05 %	85.82 %	62.07 %	0.04 s	1 core @ 2.5 Ghz (C/C++)

97	T-SSD			68.83 %	83.84 %	64.41 %	0.04	1 core @ 2.0 Ghz (C/C++)

98	DensePointPillars			68.75 %	81.53 %	62.55 %	0.03 s	GPU @ 2.5 Ghz (Python)

99	ARPNET			68.72 %	82.61 %	62.00 %	0.08 s	GPU @ 2.5 Ghz (Python + C/C++)
Y. Ye, C. Zhang and X. Hao: ARPNET: attention region proposal network for 3D object detection. Science China Information Sciences 2019.
100	PointPillars		code	68.55 %	83.79 %	61.71 %	16 ms	1080ti GPU and Intel i7 CPU
A. Lang, S. Vora, H. Caesar, L. Zhou, J. Yang and O. Beijbom: PointPillars: Fast Encoders for Object Detection from Point Clouds. CVPR 2019.
101	geo-pillars			67.88 %	80.70 %	61.40 %	0.05 s	1 core @ 2.5 Ghz (C/C++)
ERROR: Wrong syntax in BIBTEX file.
102	PCNet3D			67.69 %	81.72 %	61.51 %	0.05 s	GPU @ 3.5 Ghz (Python)

103	EPNet++			67.26 %	79.81 %	61.75 %	0.1 s	GPU @ 2.5 Ghz (Python)
Z. Liu, T. Huang, B. Li, X. Chen, X. Wang and X. Bai: EPNet++: Cascade Bi-Directional Fusion for Multi-Modal 3D Object Detection. IEEE Transactions on Pattern Analysis and Machine Intelligence 2022.
104	TANet		code	66.37 %	81.15 %	60.10 %	0.035s	GPU @ 2.5 Ghz (Python + C/C++)
Z. Liu, X. Zhao, T. Huang, R. Hu, Y. Zhou and X. Bai: TANet: Robust 3D Object Detection from Point Clouds with Triple Attention. AAAI 2020.
105	PointRGBNet			65.68 %	79.64 %	59.48 %	0.08 s	4 cores @ 2.5 Ghz (Python + C/C++)
P. Xie Desheng: Real-time Detection of 3D Objects Based on Multi-Sensor Information Fusion. Automotive Engineering 2022.
106	Fade-kd			64.57 %	78.65 %	58.66 %	0.01 s	1 core @ 2.5 Ghz (C/C++)

107	PFF3D		code	64.06 %	78.02 %	58.06 %	0.05 s	GPU @ 3.0 Ghz (Python + C/C++)
L. Wen and K. Jo: Fast and Accurate 3D Object Detection for Lidar-Camera-Based Autonomous Vehicles Using One Shared Voxel-Based Backbone. IEEE Access 2021.
108	PillarHist			63.79 %	76.08 %	57.47 %	0.01 s	1 core @ 2.5 Ghz (C/C++)

109	SeSame-pillar		code	63.61 %	75.66 %	57.48 %	N/A s	TITAN RTX @ 1.35 Ghz (Python)
H. O, C. Yang and K. Huh: SeSame: Simple, Easy 3D Object Detection with Point-Wise Semantics. Proceedings of the Asian Conference on Computer Vision (ACCV) 2024.
110	SubCNN			63.36 %	71.97 %	55.42 %	2 s	GPU @ 3.5 Ghz (Python + C/C++)
Y. Xiang, W. Choi, Y. Lin and S. Savarese: Subcategory-aware Convolutional Neural Networks for Object Proposals and Detection. IEEE Winter Conference on Applications of Computer Vision (WACV) 2017.
111	PiFeNet		code	62.62 %	77.54 %	55.66 %	0.03 s	1 core @ 2.5 Ghz (C/C++)
D. Le, H. Shi, H. Rezatofighi and J. Cai: Accurate and Real-time 3D Pedestrian Detection Using an Efficient Attentive Pillar Network. IEEE Robotics and Automation Letters 2022.
112	Pose-RCNN			62.02 %	75.74 %	53.99 %	2 s	>8 cores @ 2.5 Ghz (Python)
M. Braun, Q. Rao, Y. Wang and F. Flohr: Pose-RCNN: Joint object detection and pose estimation using 3D object proposals. Intelligent Transportation Systems (ITSC), 2016 IEEE 19th International Conference on 2016.
113	SCNet			61.11 %	77.77 %	54.82 %	0.04 s	GPU @ 3.0 Ghz (Python)
Z. Wang, H. Fu, L. Wang, L. Xiao and B. Dai: SCNet: Subdivision Coding Network for Object Detection Based on 3D Point Cloud. IEEE Access 2019.
114	AEPF			61.08 %	74.43 %	55.71 %	0.05 s	GPU @ 2.5 Ghz (Python)

115	DMF			60.85 %	71.83 %	54.58 %	0.2 s	1 core @ 2.5 Ghz (Python + C/C++)
X. J. Chen and W. Xu: Disparity-Based Multiscale Fusion Network for Transportation Detection. IEEE Transactions on Intelligent Transportation Systems 2022.
116	BirdNet+		code	59.44 %	67.52 %	54.27 %	0.11 s	Titan Xp (PyTorch)
A. Barrera, J. Beltrán, C. Guindel, J. Iglesias and F. García: BirdNet+: Two-Stage 3D Object Detection in LiDAR through a Sparsity-Invariant Bird’s Eye View. IEEE Access 2021.
117	PL++: PV-RCNN++			59.07 %	73.67 %	52.52 %	0.342 s	RTX 4060Ti (Python)

118	AVOD-FPN		code	58.70 %	69.21 %	53.47 %	0.1 s	Titan X (Pascal)
J. Ku, M. Mozifian, J. Lee, A. Harakeh and S. Waslander: Joint 3D Proposal Generation and Object Detection from View Aggregation. IROS 2018.
119	Deep3DBox			58.56 %	68.31 %	50.30 %	1.5 s	GPU @ 2.5 Ghz (C/C++)
A. Mousavian, D. Anguelov, J. Flynn and J. Kosecka: 3D Bounding Box Estimation Using Deep Learning and Geometry. CVPR 2017.
120	3DOP		code	58.45 %	72.24 %	51.91 %	3s	GPU @ 2.5 Ghz (Matlab + C/C++)
X. Chen, K. Kundu, Y. Zhu, A. Berneshawi, H. Ma, S. Fidler and R. Urtasun: 3D Object Proposals for Accurate Object Class Detection. NIPS 2015.
121	Complexer-YOLO			58.28 %	65.41 %	54.27 %	0.06 s	GPU @ 3.5 Ghz (C/C++)
M. Simon, K. Amende, A. Kraus, J. Honer, T. Samann, H. Kaulbersch, S. Milz and H. Michael Gross: Complexer-YOLO: Real-Time 3D Object Detection and Tracking on Semantic Point Clouds. The IEEE Conference on Computer Vision and Pattern Recognition (CVPR) Workshops 2019.
122	Fade			57.98 %	73.50 %	52.62 %	0.01 s	1 core @ 2.5 Ghz (C/C++)

123	DD3D		code	57.42 %	73.60 %	50.90 %	n/a s	1 core @ 2.5 Ghz (C/C++)
D. Park, R. Ambrus, V. Guizilini, J. Li and A. Gaidon: Is Pseudo-Lidar needed for Monocular 3D Object detection?. IEEE/CVF International Conference on Computer Vision (ICCV) .
124	DeepStereoOP			56.55 %	69.36 %	49.37 %	3.4 s	GPU @ 3.5 Ghz (Matlab + C/C++)
C. Pham and J. Jeon: Robust Object Proposals Re-ranking for Object Detection in Autonomous Driving Using Convolutional Neural Networks. Signal Processing: Image Communiation 2017.
125	MonoLiG		code	54.91 %	76.10 %	47.58 %	0.03 s	1 core @ 2.5 Ghz (C/C++)
A. Hekimoglu, M. Schmidt and A. Ramiro: Monocular 3D Object Detection with LiDAR Guided Semi Supervised Active Learning. 2023.
126	SeSame-voxel w/score		code	54.49 %	66.51 %	49.51 %	N/A s	GPU @ 1.5 Ghz (Python)
H. O, C. Yang and K. Huh: SeSame: Simple, Easy 3D Object Detection with Point-Wise Semantics. Proceedings of the Asian Conference on Computer Vision (ACCV) 2024.
127	Mix-Teaching		code	54.00 %	70.90 %	46.66 %	30 s	1 core @ 2.5 Ghz (C/C++)
L. Yang, X. Zhang, L. Wang, M. Zhu, C. Zhang and J. Li: Mix-Teaching: A Simple, Unified and Effective Semi-Supervised Learning Framework for Monocular 3D Object Detection. ArXiv 2022.
128	Mono3D		code	53.96 %	67.33 %	47.91 %	4.2 s	GPU @ 2.5 Ghz (Matlab + C/C++)
X. Chen, K. Kundu, Z. Zhang, H. Ma, S. Fidler and R. Urtasun: Monocular 3D Object Detection for Autonomous Driving. CVPR 2016.
129	AVOD		code	51.05 %	64.81 %	45.12 %	0.08 s	Titan X (pascal)
J. Ku, M. Mozifian, J. Lee, A. Harakeh and S. Waslander: Joint 3D Proposal Generation and Object Detection from View Aggregation. IROS 2018.
130	BirdNet+ (legacy)		code	50.94 %	69.92 %	47.01 %	0.1 s	Titan Xp (PyTorch)
A. Barrera, C. Guindel, J. Beltrán and F. García: BirdNet+: End-to-End 3D Object Detection in LiDAR Bird’s Eye View. 2020 IEEE 23rd International Conference on Intelligent Transportation Systems (ITSC) 2020.
131	FRCNN+Or		code	49.53 %	63.45 %	43.65 %	0.09 s	Titan Xp GPU
C. Guindel, D. Martin and J. Armingol: Fast Joint Object Detection and Viewpoint Estimation for Traffic Scene Understanding. IEEE Intelligent Transportation Systems Magazine 2018. C. Guindel, D. Martin and J. Armingol: Joint Object Detection and Viewpoint Estimation using CNN features. IEEE International Conference on Vehicular Electronics and Safety (ICVES) 2017.
132	MonoPSR		code	49.32 %	58.63 %	43.05 %	0.2 s	GPU @ 3.5 Ghz (Python)
J. Ku, A. Pon and S. Waslander: Monocular 3D Object Detection Leveraging Accurate Proposals and Shape Reconstruction. CVPR 2019.
133	StereoDistill			48.99 %	65.65 %	43.14 %	0.4 s	1 core @ 2.5 Ghz (Python)
Z. Liu, X. Ye, X. Tan, D. Errui, Y. Zhou and X. Bai: StereoDistill: Pick the Cream from LiDAR for Distilling Stereo-based 3D Object Detection. Proceedings of the AAAI Conference on Artificial Intelligence 2023.
134	MonoFlex			47.91 %	65.51 %	40.40 %	0.03 s	GPU @ 2.5 Ghz (Python)
Y. Zhang, J. Lu and J. Zhou: Objects are Different: Flexible Monocular 3D Object Detection. CVPR 2021.
135	HomoLoss(monoflex)		code	47.36 %	62.89 %	40.55 %	0.04 s	1 core @ 2.5 Ghz (Python)
J. Gu, B. Wu, L. Fan, J. Huang, S. Cao, Z. Xiang and X. Hua: Homography Loss for Monocular 3D Object Detection. Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) 2022.
136	MonoLSS			47.09 %	65.31 %	41.74 %	0.04 s	1 core @ 2.5 Ghz (Python)
Z. Li, J. Jia and Y. Shi: MonoLSS: Learnable Sample Selection For Monocular 3D Detection. International Conference on 3D Vision 2024.
137	MonoAFKD			46.76 %	64.39 %	41.36 %	0.03 s	1 core @ 2.5 Ghz (C/C++)

138	QD-3DT		code	46.24 %	64.64 %	40.58 %	0.03 s	GPU @ 2.5 Ghz (Python)
H. Hu, Y. Yang, T. Fischer, F. Yu, T. Darrell and M. Sun: Monocular Quasi-Dense 3D Object Tracking. ArXiv:2103.07351 2021.
139	DSGN++		code	45.94 %	57.93 %	41.93 %	0.2 s	GeForce RTX 2080Ti
Y. Chen, S. Huang, S. Liu, B. Yu and J. Jia: DSGN++: Exploiting Visual-Spatial Relation for Stereo-Based 3D Detectors. IEEE Transactions on Pattern Analysis and Machine Intelligence 2022.
140	MonoDDE			45.58 %	63.91 %	39.29 %	0.04 s	1 core @ 2.5 Ghz (Python)
Z. Li, Z. Qu, Y. Zhou, J. Liu, H. Wang and L. Jiang: Diversity Matters: Fully Exploiting Depth Clues for Reliable Monocular 3D Object Detection. CVPR 2022.
141	LPCG-Monoflex		code	45.24 %	63.07 %	39.28 %	0.03 s	1 core @ 2.5 Ghz (C/C++)
L. Peng, F. Liu, Z. Yu, S. Yan, D. Deng, Z. Yang, H. Liu and D. Cai: Lidar Point Cloud Guided Monocular 3D Object Detection. ECCV 2022.
142	MonoUNI		code	45.21 %	62.21 %	38.28 %	0.04 s	1 core @ 2.5 Ghz (Python)
J. Jia, Z. Li and Y. Shi: MonoUNI: A Unified Vehicle and Infrastructure-side Monocular 3D Object Detection Network with Sufficient Depth Clues. Thirty-seventh Conference on Neural Information Processing Systems 2023.
143	MMLAB LIGA-Stereo		code	45.13 %	63.89 %	39.23 %	0.4 s	1 core @ 2.5 Ghz (Python + C/C++)
X. Guo, S. Shi, X. Wang and H. Li: LIGA-Stereo: Learning LiDAR Geometry Aware Representations for Stereo-based 3D Detector. Proceedings of the IEEE/CVF International Conference on Computer Vision (ICCV) 2021.
144	monodle		code	45.12 %	61.84 %	37.95 %	0.04 s	GPU @ 2.5 Ghz (Python)
X. Ma, Y. Zhang, D. Xu, D. Zhou, S. Yi, H. Li and W. Ouyang: Delving into Localization Errors for Monocular 3D Object Detection. CVPR 2021 .
145	BirdNet			45.03 %	62.69 %	41.88 %	0.11 s	Titan Xp (Caffe)
J. Beltrán, C. Guindel, F. Moreno, D. Cruzado, F. García and A. Escalera: BirdNet: A 3D Object Detection Framework from LiDAR Information. 2018 21st International Conference on Intelligent Transportation Systems (ITSC) 2018.
146	MonOAPC			44.74 %	60.40 %	38.01 %	0035 s	1 core @ 2.5 Ghz (Python)
H. Yao, J. Chen, Z. Wang, X. Wang, P. Han, X. Chai and Y. Qiu: Occlusion-Aware Plane-Constraints for Monocular 3D Object Detection. IEEE Transactions on Intelligent Transportation Systems 2023.
147	AM			44.03 %	59.88 %	39.11 %	0.04 s	1 core @ 2.5 Ghz (C/C++)

148	SparsePool		code	43.50 %	59.77 %	39.36 %	0.13 s	8 cores @ 2.5 Ghz (Python)
Z. Wang, W. Zhan and M. Tomizuka: Fusing bird view lidar point cloud and front view camera image for deep object detection. arXiv preprint arXiv:1711.06703 2017.
149	MonoDTR			42.45 %	56.40 %	36.32 %	0.04 s	1 core @ 2.5 Ghz (C/C++)
K. Huang, T. Wu, H. Su and W. Hsu: MonoDTR: Monocular 3D Object Detection with Depth-Aware Transformer. CVPR 2022.
150	UniCuboid			41.71 %	60.19 %	36.43 %	0.1 s	1 core @ 2.5 Ghz (C/C++)

151	sensekitti		code	41.14 %	47.48 %	35.07 %	4.5 s	GPU @ 2.5 Ghz (Python + C/C++)
B. Yang, J. Yan, Z. Lei and S. Li: Craft Objects from Images. CVPR 2016.
152	CG-Stereo			40.64 %	60.24 %	35.55 %	0.57 s	GeForce RTX 2080 Ti
C. Li, J. Ku and S. Waslander: Confidence Guided Stereo 3D Object Detection with Split Depth Estimation. IROS 2020.
153	MonoRCNN++		code	39.84 %	56.32 %	34.82 %	0.07 s	GPU @ 2.5 Ghz (Python)
X. Shi, Z. Chen and T. Kim: Multivariate Probabilistic Monocular 3D Object Detection. WACV 2023.
154	MonoPair			39.47 %	53.36 %	33.95 %	0.06 s	GPU @ 2.5 Ghz (Python + C/C++)
Y. Chen, L. Tai, K. Sun and M. Li: MonoPair: Monocular 3D Object Detection Using Pairwise Spatial Relationships. The IEEE Conference on Computer Vision and Pattern Recognition (CVPR) 2020.
155	CMKD		code	38.70 %	56.46 %	34.00 %	0.1 s	1 core @ 2.5 Ghz (C/C++)
Y. Hong, H. Dai and Y. Ding: Cross-Modality Knowledge Distillation Network for Monocular 3D Object Detection. ECCV 2022.
156	DEVIANT		code	38.46 %	57.64 %	32.76 %	0.04 s	1 GPU (Python)
A. Kumar, G. Brazil, E. Corona, A. Parchami and X. Liu: DEVIANT: Depth EquiVarIAnt NeTwork for Monocular 3D Object Detection. European Conference on Computer Vision (ECCV) 2022.
157	monospb			37.94 %	54.36 %	32.68 %	0.01 s	1 core @ 2.5 Ghz (Python)

158	SAKD-MR-Res18			36.74 %	52.25 %	32.07 %	0.03 s	1 core @ 2.5 Ghz (C/C++)

159	Disp R-CNN (velo)		code	35.93 %	52.35 %	31.09 %	0.387 s	GPU @ 2.5 Ghz (Python + C/C++)
J. Sun, L. Chen, Y. Xie, S. Zhang, Q. Jiang, X. Zhou and H. Bao: Disp R-CNN: Stereo 3D Object Detection via Shape Prior Guided Instance Disparity Estimation. CVPR 2020.
160	Disp R-CNN		code	35.92 %	52.37 %	31.08 %	0.387 s	GPU @ 2.5 Ghz (Python + C/C++)
J. Sun, L. Chen, Y. Xie, S. Zhang, Q. Jiang, X. Zhou and H. Bao: Disp R-CNN: Stereo 3D Object Detection via Shape Prior Guided Instance Disparity Estimation. CVPR 2020.
161	MonoCoP			35.67 %	51.99 %	32.11 %	0.01 s	1 core @ 2.5 Ghz (C/C++)

162	GUPNet		code	35.03 %	55.03 %	31.18 %	NA s	1 core @ 2.5 Ghz (Python + C/C++)
Y. Lu, X. Ma, L. Yang, T. Zhang, Y. Liu, Q. Chu, J. Yan and W. Ouyang: Geometry Uncertainty Projection Network for Monocular 3D Object Detection. arXiv preprint arXiv:2107.13774 2021.
163	Shift R-CNN (mono)		code	34.77 %	51.95 %	31.10 %	0.25 s	GPU @ 1.5 Ghz (Python)
A. Naiden, V. Paunescu, G. Kim, B. Jeon and M. Leordeanu: Shift R-CNN: Deep Monocular 3D Object Detection With Closed-form Geometric Constraints. ICIP 2019.
164	SparsePool		code	34.56 %	43.33 %	31.09 %	0.13 s	8 cores @ 2.5 Ghz (Python)
Z. Wang, W. Zhan and M. Tomizuka: Fusing bird view lidar point cloud and front view camera image for deep object detection. arXiv preprint arXiv:1711.06703 2017.
165	MonoRUn		code	34.36 %	49.04 %	30.22 %	0.07 s	GPU @ 2.5 Ghz (Python + C/C++)
H. Chen, Y. Huang, W. Tian, Z. Gao and L. Xiong: MonoRUn: Monocular 3D Object Detection by Reconstruction and Uncertainty Propagation. IEEE Conference on Computer Vision and Pattern Recognition (CVPR) 2021.
166	SPG_mini		code	34.28 %	36.23 %	32.09 %	0.09 s	GPU @ 2.5 Ghz (Python)
Q. Xu, Y. Zhou, W. Wang, C. Qi and D. Anguelov: SPG: Unsupervised Domain Adaptation for 3D Object Detection via Semantic Point Generation. Proceedings of the IEEE conference on computer vision and pattern recognition (ICCV) 2021.
167	MonoMH		code	34.13 %	51.00 %	29.36 %	0.04 s	1 core @ 2.5 Ghz (Python)

168	BtcDet		code	33.94 %	35.79 %	31.90 %	0.09 s	GPU @ 2.5 Ghz (Python + C/C++)
Q. Xu, Y. Zhong and U. Neumann: Behind the Curtain: Learning Occluded Shapes for 3D Object Detection. Proceedings of the AAAI Conference on Artificial Intelligence 2022.
169	Plane-Constraints		code	32.87 %	48.36 %	28.52 %	0.05 s	4 cores @ 3.0 Ghz (Python)
H. Yao, J. Chen, Z. Wang, X. Wang, X. Chai, Y. Qiu and P. Han: Vertex points are not enough: Monocular 3D object detection via intra-and inter-plane constraints. Neural Networks 2023.
170	Point-GNN		code	32.37 %	36.29 %	29.81 %	0.6 s	GPU @ 2.5 Ghz (Python)
W. Shi and R. Rajkumar: Point-GNN: Graph Neural Network for 3D Object Detection in a Point Cloud. CVPR 2020.
171	MonoEF			32.19 %	43.70 %	27.93 %	0.03 s	1 core @ 2.5 Ghz (Python)
Y. Zhou, Y. He, H. Zhu, C. Wang, H. Li and Q. Jiang: Monocular 3D Object Detection: An Extrinsic Parameter Free Approach. Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) 2021.
172	D4LCN		code	31.70 %	48.03 %	26.99 %	0.2 s	GPU @ 2.5 Ghz (Python + C/C++)
M. Ding, Y. Huo, H. Yi, Z. Wang, J. Shi, Z. Lu and P. Luo: Learning Depth-Guided Convolutions for Monocular 3D Object Detection. CVPR 2020.
173	OPA-3D		code	31.64 %	45.97 %	27.92 %	0.04 s	1 core @ 3.5 Ghz (Python)
Y. Su, Y. Di, G. Zhai, F. Manhardt, J. Rambach, B. Busam, D. Stricker and F. Tombari: OPA-3D: Occlusion-Aware Pixel-Wise Aggregation for Monocular 3D Object Detection. IEEE Robotics and Automation Letters 2023.
174	M3D-RPN		code	31.09 %	48.11 %	26.10 %	0.16 s	GPU @ 1.5 Ghz (Python)
G. Brazil and X. Liu: M3D-RPN: Monocular 3D Region Proposal Network for Object Detection . ICCV 2019 .
175	Aug3D-RPN			30.01 %	42.60 %	24.74 %	0.08 s	1 core @ 2.5 Ghz (C/C++)
C. He, J. Huang, X. Hua and L. Zhang: Aug3D-RPN: Improving Monocular 3D Object Detection by Synthetic Images with Virtual Depth. 2021.
176	DDMP-3D			29.53 %	46.42 %	25.91 %	0.18 s	1 core @ 2.5 Ghz (Python)
L. Wang, L. Du, X. Ye, Y. Fu, G. Guo, X. Xue, J. Feng and L. Zhang: Depth-conditioned Dynamic Message Propagation for Monocular 3D Object Detection. CVPR 2020.
177	PS-fld		code	27.99 %	41.21 %	24.75 %	0.25 s	1 core @ 2.5 Ghz (C/C++)
Y. Chen, H. Dai and Y. Ding: Pseudo-Stereo for Monocular 3D Object Detection in Autonomous Driving. Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) 2022.
178	SS3D			27.79 %	42.95 %	24.26 %	48 ms	Tesla V100 (Python)
E. Jörgensen, C. Zach and F. Kahl: Monocular 3D Object Detection and Box Fitting Trained End-to-End Using Intersection-over-Union Loss. CoRR 2019.
179	CMAN			27.63 %	42.58 %	23.14 %	0.15 s	1 core @ 2.5 Ghz (Python)
C. Yuanzhouhan Cao: CMAN: Leaning Global Structure Correlation for Monocular 3D Object Detection. IEEE Trans. Intell. Transport. Syst. 2022.
180	AMNet+DDAD15M		code	27.07 %	40.75 %	23.54 %	0.03 s	1 core @ 2.5 Ghz (C/C++)
H. Pan, Y. Jia, J. Wang and W. Sun: MonoAMNet: Three-Stage Real-Time Monocular 3D Object Detection With Adaptive Methods. IEEE Transactions on Intelligent Transportation Systems 2025.
181	MonoFRD			26.87 %	39.05 %	24.09 %	0.01 s	1 core @ 2.5 Ghz (C/C++)
Z. Gong, Y. Zhao, F. Zhang, G. Gui, B. Chen, L. Yu, H. Wang, C. Yang and W. Gui: Color intuitive feature guided depth-height fusion and volume rendering for monocular 3D object detection. IEEE Transactions on Intelligent Vehicles(Major Revison) 2024.
182	AMNet		code	24.99 %	37.82 %	21.68 %	0.03 s	GPU @ 1.0 Ghz (Python)
H. Pan, Y. Jia, J. Wang and W. Sun: MonoAMNet: Three-Stage Real-Time Monocular 3D Object Detection With Adaptive Methods. IEEE Transactions on Intelligent Transportation Systems 2025.
183	DFR-Net			24.85 %	38.60 %	21.86 %	0.18 s	1080 Ti (Pytorch)
Z. Zou, X. Ye, L. Du, X. Cheng, X. Tan, L. Zhang, J. Feng, X. Xue and E. Ding: The devil is in the task: Exploiting reciprocal appearance-localization features for monocular 3d object detection . ICCV 2021.
184	Cube R-CNN		code	23.98 %	29.00 %	21.67 %	0.05 s	GPU @ 2.5 Ghz (Python)
G. Brazil, A. Kumar, J. Straub, N. Ravi, J. Johnson and G. Gkioxari: Omni3D: A Large Benchmark and Model for 3D Object Detection in the Wild. CVPR 2023.
185	mdab			23.84 %	35.49 %	21.15 %	0.02 s	1 core @ 2.5 Ghz (Python)

186	mdab			23.84 %	35.49 %	21.15 %	22 s	1 core @ 2.5 Ghz (C/C++)

187	SeSame-pillar w/scor		code	21.79 %	19.53 %	20.12 %	N/A s	1 core @ 2.5 Ghz (C/C++)
H. O, C. Yang and K. Huh: SeSame: Simple, Easy 3D Object Detection with Point-Wise Semantics. Proceedings of the Asian Conference on Computer Vision (ACCV) 2024.
188	CPD++(unsupervised)		code	21.26 %	28.13 %	17.94 %	0.1 s	GPU @ >3.5 Ghz (Python)

189	DSGN		code	20.28 %	29.76 %	19.13 %	0.67 s	NVIDIA Tesla V100
Y. Chen, S. Liu, X. Shen and J. Jia: DSGN: Deep Stereo Geometry Network for 3D Object Detection. CVPR 2020.
190	MonoNeRD		code	20.13 %	30.64 %	18.19 %	na s	1 core @ 2.5 Ghz (Python)
J. Xu, L. Peng, H. Cheng, H. Li, W. Qian, K. Li, W. Wang and D. Cai: MonoNeRD: NeRF-like Representations for Monocular 3D Object Detection. ICCV 2023.
191	CaDDN		code	19.96 %	30.35 %	17.38 %	0.63 s	GPU @ 2.5 Ghz (Python)
C. Reading, A. Harakeh, J. Chae and S. Waslander: Categorical Depth Distribution Network for Monocular 3D Object Detection. CVPR 2021.
192	PS-SVDM			19.50 %	29.75 %	17.08 %	1 s	1 core @ 2.5 Ghz (Python)
Y. Shi: SVDM: Single-View Diffusion Model for Pseudo-Stereo 3D Object Detection. arXiv preprint arXiv:2307.02270 2023.
193	PS-SVDM			19.46 %	30.41 %	16.67 %	1 s	1 core @ 2.5 Ghz (Python)
Y. Shi: SVDM: Single-View Diffusion Model for Pseudo-Stereo 3D Object Detection. arXiv preprint arXiv:2307.02270 2023.
194	LSVM-MDPM-sv			19.15 %	26.05 %	18.02 %	10 s	4 cores @ 3.0 Ghz (C/C++)
P. Felzenszwalb, R. Girshick, D. McAllester and D. Ramanan: Object Detection with Discriminatively Trained Part-Based Models. PAMI 2010. A. Geiger, C. Wojek and R. Urtasun: Joint 3D Estimation of Objects and Scene Layout. NIPS 2011.
195	PGD-FCOS3D		code	19.10 %	31.75 %	16.59 %	0.03 s	1 core @ 2.5 Ghz (C/C++)
T. Wang, X. Zhu, J. Pang and D. Lin: Probabilistic and Geometric Depth: Detecting Objects in Perspective. Conference on Robot Learning (CoRL) 2021.
196	OC Stereo		code	18.99 %	29.07 %	16.40 %	0.35 s	1 core @ 2.5 Ghz (Python + C/C++)
A. Pon, J. Ku, C. Li and S. Waslander: Object-Centric Stereo Matching for 3D Object Detection. ICRA 2020.
197	DPM-VOC+VP			18.92 %	27.97 %	17.43 %	8 s	1 core @ 2.5 Ghz (C/C++)
B. Pepik, M. Stark, P. Gehler and B. Schiele: Multi-view and 3D Deformable Part Models. IEEE Transactions on Pattern Analysis and Machine Intelligence (TPAMI) 2015.
198	CIE			17.52 %	24.39 %	15.84 %	0.1 s	1 core @ 2.5 Ghz (C/C++)
Anonymities: Consistency of Implicit and Explicit Features Matters for Monocular 3D Object Detection. arXiv preprint arXiv:2207.07933 2022.
199	SGM3D		code	16.50 %	25.51 %	15.09 %	0.03 s	1 core @ 2.5 Ghz (C/C++)
Z. Zhou, L. Du, X. Ye, Z. Zou, X. Tan, L. Zhang, X. Xue and J. Feng: SGM3D: Stereo Guided Monocular 3D Object Detection. RA-L 2022.
200	RT3D-GMP			16.18 %	23.91 %	14.23 %	0.06 s	GPU @ 2.5 Ghz (Python + C/C++)
H. Königshof and C. Stiller: Learning-Based Shape Estimation with Grid Map Patches for Realtime 3D Object Detection for Automated Driving. 2020 IEEE 23rd International Conference on Intelligent Transportation Systems (ITSC) 2020.
201	RefinedMPL			16.02 %	26.54 %	13.20 %	0.15 s	GPU @ 2.5 Ghz (Python + C/C++)
J. Vianney, S. Aich and B. Liu: RefinedMPL: Refined Monocular PseudoLiDAR for 3D Object Detection in Autonomous Driving. arXiv preprint arXiv:1911.09712 2019.
202	FMF-occlusion-net			15.24 %	23.82 %	13.84 %	0.16 s	1 core @ 2.5 Ghz (Python + C/C++)
H. Liu, H. Liu, Y. Wang, F. Sun and W. Huang: Fine-grained Multi-level Fusion for Anti- occlusion Monocular 3D Object Detection. IEEE Transactions on Image Processing 2022.
203	DPM-C8B1			14.64 %	23.93 %	13.09 %	15 s	4 cores @ 2.5 Ghz (Matlab + C/C++)
J. Yebes, L. Bergasa and M. García-Garrido: Visual Object Recognition with 3D-Aware Features in KITTI Urban Scenes. Sensors 2015. J. Yebes, L. Bergasa, R. Arroyo and A. Lázaro: Supervised learning and evaluation of KITTI's cars detector with DPM. IV 2014.
204	LLW			11.65 %	18.38 %	10.59 %	0.1 s	1 core @ 2.5 Ghz (C/C++)

205	SeSame-point w/score		code	10.17 %	12.39 %	9.31 %	N/A s	1 core @ 1.5 Ghz (Python)
H. O, C. Yang and K. Huh: SeSame: Simple, Easy 3D Object Detection with Point-Wise Semantics. Proceedings of the Asian Conference on Computer Vision (ACCV) 2024.
206	ESGN			7.73 %	12.50 %	6.80 %	0.06 s	GPU @ 2.5 Ghz (Python + C/C++)
A. Gao, Y. Pang, J. Nie, Z. Shao, J. Cao, Y. Guo and X. Li: ESGN: Efficient Stereo Geometry Network for Fast 3D Object Detection. IEEE Transactions on Circuits and Systems for Video Technology 2022.
207	CPD(unsupervised)		code	4.97 %	7.07 %	4.49 %	0.1 s	GPU @ >3.5 Ghz (Python + C/C++)
H. Wu, S. Zhao, X. Huang, C. Wen, X. Li and C. Wang: Commonsense Prototype for Outdoor Unsupervised 3D Object Detection. CVPR 2024.
208	RT3DStereo			3.88 %	5.46 %	3.54 %	0.08 s	GPU @ 2.5 Ghz (C/C++)
H. Königshof, N. Salscheider and C. Stiller: Realtime 3D Object Detection for Automated Driving Using Stereo Vision and Semantic Information. Proc. IEEE Intl. Conf. Intelligent Transportation Systems 2019.
209	GATE3D		code	0.26 %	0.00 %	0.44 %	0.01 s	1 core @ 2.5 Ghz (C/C++)

210	f3sd		code	0.01 %	0.02 %	0.02 %	1.67 s	1 core @ 2.5 Ghz (C/C++)

211	DA3D+KM3D+v2-99		code	0.00 %	0.00 %	0.00 %	0.120s	GPU @ 2.5 Ghz (Python)
Y. Jia, J. Wang, H. Pan and W. Sun: Enhancing Monocular 3-D Object Detection Through Data Augmentation Strategies. IEEE Transactions on Instrumentation and Measurement 2024.
212	DA3D+KM3D		code	0.00 %	0.00 %	0.00 %	0.02 s	GPU @ 2.5 Ghz (Python)
Y. Jia, J. Wang, H. Pan and W. Sun: Enhancing Monocular 3-D Object Detection Through Data Augmentation Strategies. IEEE Transactions on Instrumentation and Measurement 2024.
213	DA3D		code	0.00 %	0.00 %	0.00 %	0.03 s	1 core @ 2.5 Ghz (Python)
Y. Jia, J. Wang, H. Pan and W. Sun: Enhancing Monocular 3-D Object Detection Through Data Augmentation Strategies. IEEE Transactions on Instrumentation and Measurement 2024.

Related Datasets

CERV Vehicle Lights Dataset: Annotations of vehicle lights for a subset of the object detection benchmark.
PASCAL3D+: Augments 12 rigid object classes of PASCAL VOC 2012 with 3D annotations.
The PASCAL Visual Object Classes Challenges: Dataset and benchmarks for object class recognition.
TME Motorway Dataset: 28 video sequences with vehicle annotations captured from VisLab's BRAiVE vehicle.
LabelMe: Online annotation tool to build image databases for computer vision research.
MIT Street Scenes: Street-side images with labels for 9 object categories (including cars, pedestrians, buildings, trees).
Daimler Pedestrian Datasets: Datasets focusing on pedestrian detection for autonomous driving.
Caltech Pedestrian Detection Benchmark: 10 hours of video with 350.000 annotated pedestrian bounding boxes.
Robust Multi-Person Tracking from Mobile Platforms: Videos with annotated pedestrians captured from a stroller.

Citation

When using this dataset in your research, we will be happy if you cite us:
@inproceedings{Geiger2012CVPR,
author = {Andreas Geiger and Philip Lenz and Raquel Urtasun},
title = {Are we ready for Autonomous Driving? The KITTI Vision Benchmark Suite},
booktitle = {Conference on Computer Vision and Pattern Recognition (CVPR)},
year = {2012}
}

The KITTI Vision Benchmark Suite

A project of Karlsruhe Institute of Technology
and Toyota Technological Institute at Chicago

Object Detection Evaluation 2012

Car

Pedestrian

Cyclist

Object Detection and Orientation Estimation Evaluation

Cars

Pedestrians

Cyclists

Related Datasets

Citation

The KITTI Vision Benchmark Suite

A project of Karlsruhe Institute of Technologyand Toyota Technological Institute at Chicago

Object Detection Evaluation 2012

Car

Pedestrian

Cyclist

Object Detection and Orientation Estimation Evaluation

Cars

Pedestrians

Cyclists

Related Datasets

Citation

A project of Karlsruhe Institute of Technology
and Toyota Technological Institute at Chicago