3 ms·
“Solve vision” is a bit ill-posed. If you can accurately determine the 3D geometry of the scene(e.g. with LiDAR), the 3D object detection task becomes much eas
by upbeat_general 4y ago
“Solve vision” is a bit ill-posed.
If you can accurately determine the 3D geometry of the scene(e.g. with LiDAR), the 3D object detection task becomes much easier.
That being said, most tasks for self-driving such as object detection can be robustly “solved” by LiDAR-only (to the extent that the important actors will be recognized) but adding in cameras obviously helps to distinguish between some classes.
Trying to do camera only (specifically monocular) runs into the issue of no 3D ground truth meaning it’s a lot more likely to accidentally not detect something in frame (say a white truck).
That’s why you can have LiDAR and partially-“solved” vision but need fully solved vision if it’s the only input.