All
Articles 173,164Blog Posts 163,816Tech Tutorials 46,199Research Papers 33,872News 21,839
⚡ AI Lessons
Reddit r/MachineLearning
👁️ Computer Vision
⚡ AI Lesson
4w ago
[R], Need some best model suggestions for Face Detection,Face Recognition,Body Detection and Body identification. [R]
need those for analysing movies. example let's say I have to find the screentime of the actor over the whole runtime of the movie and i need to do it for the
Reddit r/MachineLearning
👁️ Computer Vision
⚡ AI Lesson
1mo ago
GPT-5.5 Scores 10.6% on ActiveVision, Humans Hit 96.1% [R]
The interesting finding from a new [arXiv paper]( https://arxiv.org/abs/2607.16165 ) isn't that a frontier vision model failed a new benchmark, that happens wee
![Masked depth modeling with sensor-validity masking: reports best RMSE on 7 of 8 masked/sparse depth benchmarks, plus a controlled encoder-init study[R]](https://preview.redd.it/jfsxqgnx4sbh1.png?width=140&height=41&auto=webp&s=c9d6b3a3d3ceffc8d77a713d0c41c4011d091a2f)
Reddit r/MachineLearning
👁️ Computer Vision
⚡ AI Lesson
1mo ago
Masked depth modeling with sensor-validity masking: reports best RMSE on 7 of 8 masked/sparse depth benchmarks, plus a controlled encoder-init study[R]
<img src="https://preview.redd.it/jfsxqgnx4sbh1.png?width=140&height=41&auto=webp&s=c9d6b3a3d3ceffc8d77a713d0c41c4011d091a2f" alt="Masked depth mode
![LingBot-Vision: masked boundary modeling for self-supervised pretraining (0.296 NYUv2 linear-probe RMSE at 1.1B vs 0.309 for DINOv3-7B, trails on ImageNet); weights in 4 sizes[R]](https://preview.redd.it/ha08vg49bnbh1.png?width=140&height=78&auto=webp&s=cbd1e4aed6c0571b7f0acee245c21543fe356719)
Reddit r/MachineLearning
👁️ Computer Vision
⚡ AI Lesson
1mo ago
LingBot-Vision: masked boundary modeling for self-supervised pretraining (0.296 NYUv2 linear-probe RMSE at 1.1B vs 0.309 for DINOv3-7B, trails on ImageNet); weights in 4 sizes[R]
<img src="https://preview.redd.it/ha08vg49bnbh1.png?width=140&height=78&auto=webp&s=cbd1e4aed6c0571b7f0acee245c21543fe356719" alt="LingBot-Vision: m
Reddit r/MachineLearning
👁️ Computer Vision
⚡ AI Lesson
2mo ago
Showcase: geolocating a dashcam video without GPS, only from the footage [P]
Sharing a project I have been working on called Third Eye. It does visual geolocation. Given a video, it figures out where it was filmed using only the image co
Reddit r/MachineLearning
👁️ Computer Vision
⚡ AI Lesson
2mo ago
ECCV 2026 camera-ready deadline: June 27 or June 30? [D]
In the recent Springer/Meteor email, it says: The deadline for the upload of the camera-ready manuscripts and source files is 30 June. This is a hard deadline a
Reddit r/MachineLearning
👁️ Computer Vision
⚡ AI Lesson
2mo ago
[ECCV 2026] Paper Decision Appeals Discussion [D]
With the release of meta-reviews, ECCV sent out a google form for dissatisfied authors to submit an appeal for the following reasons: Policy errors, e.g., revie
Reddit r/MachineLearning
👁️ Computer Vision
⚡ AI Lesson
2mo ago
Any idea if AAAI will be harsh on computer vision paper as last year? [R]
Hello everyone, I have a computer vision paper ready for submission, a coauthor have suggested submitting it to AAAI. However last year computer vision papers h
![Browse CVPR 2026 papers on PapersWithCode [P]](https://preview.redd.it/se5nr2z7tt4h1.png?width=140&height=91&auto=webp&s=dc5f025b3936b18a0af5d2fba8f15a0112bdda74)
Reddit r/MachineLearning
👁️ Computer Vision
⚡ AI Lesson
3mo ago
Browse CVPR 2026 papers on PapersWithCode [P]
<a href="https://preview.redd.it/se5nr2z7tt4h1.png?width=3046&format=
Reddit r/MachineLearning
👁️ Computer Vision
⚡ AI Lesson
3mo ago
Query about non-archival workshop at CVPR-2026 [R]
My paper was recently accepted to a workshop at CVPR-2026 as non-archival acceptance. Is it mandatory for me to register to the conference as I won't be able to
Reddit r/MachineLearning
👁️ Computer Vision
⚡ AI Lesson
3mo ago
A new dataset with more that 100M hi-quality, curated images, with captions and meta data! [P]
Hello everyone. The new dataset is named MONET, is Apache 2.0 and available on HF: https://huggingface.co/datasets/jasperai/monet MONET is open, Apache 2.0-lice
![Per-pixel bounding-box regression + DBSCAN for handwritten word detection - visual walkthrough of WordDetectorNet [P]](https://preview.redd.it/qnfoh3sqjx2h1.png?width=140&height=94&auto=webp&s=e72cb3f3e061a1362a9bd5111d9e919341d48acb)
Reddit r/MachineLearning
👁️ Computer Vision
⚡ AI Lesson
3mo ago
Per-pixel bounding-box regression + DBSCAN for handwritten word detection - visual walkthrough of WordDetectorNet [P]
<img src="https://preview.redd.it/qnfoh3sqjx2h1.png?width=140&height=94&auto=webp&s=e72cb3f3e061a1362a9bd5111d9e919341d48acb" alt="Per-pixel boundin
Reddit r/MachineLearning
👁️ Computer Vision
⚡ AI Lesson
4mo ago
What should i do to have a good OD model?[P]
I’m tired of training a lot of models and trying different datasets but still my model is trash and can’t detect clearly it sometimes has mAP50 pf 80% but it is
![Mandatory In-Person Presentation in CVPR 2026 [D]](https://preview.redd.it/z5stwi8b9zug1.png?width=140&height=16&auto=webp&s=b6f43a18f31d7ae1eef9cc885f3fb97719c1015a)
Reddit r/MachineLearning
👁️ Computer Vision
⚡ AI Lesson
4mo ago
Mandatory In-Person Presentation in CVPR 2026 [D]
In the recent mail from CVPR PC about oral and poster decision
DeepCamp AI