I’ve launched 2 new course websites for my Introduction to #ComputerVision and #DeepLearning in Computer Vision courses.
🆕 Gaussian Splatting
🆕 Flow Matching
🆕 Sparse Autoencoders
🆕 Lecture Notes (work in progress)
The videos don’t include voiceovers, but that’s coming in a future revision 🤞
Share your feedback in the thread.
arXiv has updated our policy on rate limiting for all submitters.
This update was made to fairly distribute moderator time & support the arXiv community of staff, volunteers, readers & authors.
Please read our announcement to learn more: blog.arxiv.org/2026/10/01/up…
ALT arXiv logo in white over a whiteboard with math symbols.
Shots fired?
Apple plans to introduce new privacy controls for Mac users, warning of the growing risks around granting broad data access to third-party software, including AI agents bloomberg.com/news/articles/…
(1/8)
We are excited to share our new model ARROW 🏹 !
ARROW brings D4RT-style decoding to broader input types, including multi-view videos and unordered collections of images.
📜arXiv: arxiv.org/abs/2610.01314
🌐Project page: vision.rwth-aachen.de/arrow
ChatGPT can now do motion design in After Effects.
Introducing Higgsfield AI Motion Designer.
Our ChatGPT plugin understands animation principles, writes expressions, and retains context in your After Effects projects.
Try Higgsfield’s ChatGPT plugin now in After Effects.
Joshaghani et al., "What Builds the Scene? Luminance Dominates Geometry Formation in 3D Gaussian Splatting"
This work reminded me of something that I completely forgot. We used to use grayscale images because that was all we needed to infer geometry and shapes. It still holds!
Feed-forward models distort geometry even when given camera poses. In paper to appear in #NeurIPS, we recast multi-view stereo as sequence-to-sequence: a camera-aware transformer with a unified global cost volume predicts geometry for all views jointly.
arxiv.org/abs/2609.24850