It consists of 6 depth related tasks to assess response to visual cues.
They test 20 pretrained models.
As expected DAv2, DUSt3R, DINOv2 do well, but SigLIP is not bad
danier97.github.io/depthcues/
It consists of 6 depth related tasks to assess response to visual cues.
They test 20 pretrained models.
As expected DAv2, DUSt3R, DINOv2 do well, but SigLIP is not bad
danier97.github.io/depthcues/
#ColorADay #BlueTue #GreatSmokyMountains #AtmosphericPerspective #DepthCues #Knowledge #SensoryArt #EastCoastKin
#ColorADay #BlueTue #GreatSmokyMountains #AtmosphericPerspective #DepthCues #Knowledge #SensoryArt #EastCoastKin
Do automated monocular depth estimation methods use similar visual cues to humans?
To learn more, stop by poster #405 in the evening session (17:00 to 19:00) today at #CVPR2025.
Do automated monocular depth estimation methods use similar visual cues to humans?
To learn more, stop by poster #405 in the evening session (17:00 to 19:00) today at #CVPR2025.
DepthCues: Evaluating Monocular Depth Perception in Large Vision Models
https://arxiv.org/abs/2411.17385
DepthCues: Evaluating Monocular Depth Perception in Large Vision Models
https://arxiv.org/abs/2411.17385
arxiv.org/abs/2411.17385
Project page:
danier97.github.io/depthcues/
Work led by Duolikun Danier:
danier97.github.io
arxiv.org/abs/2411.17385
Project page:
danier97.github.io/depthcues/
Work led by Duolikun Danier:
danier97.github.io