In your paper, it is reported that the method achieved 7.28% meanIoU on semanticKITTI dataset. How should this result be interpreted in practice? In particular, is this performance considered competitive for the open-vocabulary or zero-shot setting, and what are the main factors that explain the relatively low absolute mIoU? It would also be helpful to understand which baseline methods provide the most meaningful comparison and what this result demonstrates about the method’s strengths and limitations.
In your paper, it is reported that the method achieved 7.28% meanIoU on semanticKITTI dataset. How should this result be interpreted in practice? In particular, is this performance considered competitive for the open-vocabulary or zero-shot setting, and what are the main factors that explain the relatively low absolute mIoU? It would also be helpful to understand which baseline methods provide the most meaningful comparison and what this result demonstrates about the method’s strengths and limitations.