Table 8

Ablation studies of the backbone structure adoption on the mod level of the KITTI val set and all the results are evaluated with average precision and calculated by 40 recall positions. Also, we provide FPS for the comparison of the model efficiency.

MethodsCar3D Moderate PedestrianCyclistFPS
Voxel-RCNN, [6]84.8759.8773.2330
FocalConv*, [4]85.2660.0473.0819
Ours85.3460.3074.3329

Note: The above results are reproduced by the publicly release model ([28]).

*

:Note that we only adopt the attention backbone structure while the remaining part is the same as the baseline model.

or Create an Account

Close subscription notice
Close access options