Introducing D4RT: A unified AI model for 4D scene reconstruction and tracking across space and time. 🎯 Catch the demo with Skanda Koppula at 12 pm at our #CVPR2026 Google booth kiosk! d4rt-paper.github.io@GoogleDeepMind
🚀 Exciting news! We’re introducing VGG-T³: a scalable model for offline feed-forward 3D reconstruction that finally tackles the "quadratic bottleneck."
Ever wanted to have VGGT reconstruct a 1,000-image scene in seconds instead of 10 minutes and use it for visual localization?
Then, they process the images with pixel-aligned generative models to synthesize versions in thermal, night, and depth formats. This step transforms the established matches into multiple modalities.
The result is a massive dataset of 800 million synthetic image pairs. Dump these into SOTA marching models, retrain, and boom, you've got the most universal image matcher out there! 🚀
MatchAnything is genuinely an insane framework, generalizing to many modalities well beyond its training set. Simply the confirmation that scaling data is always one way to go :)
Paper: arxiv.org/abs/2501.07556
Website: zju3dv.github.io/MatchAnything/
I packaged up the "autoresearch" project into a new self-contained minimal repo if people would like to play over the weekend. It's basically nanochat LLM training core stripped down to a single-GPU, one file version of ~630 lines of code, then:
- the human iterates on the
to improve fine-tuning data efficiency, replay generic pre-training data
not only does this reduce forgetting, it actually improves performance on the fine-tuning domain! especially when fine-tuning data is scarce in pre-training (w/ @percyliang)
Forgot to share this: latest COLMAP (v3.14) and pycolmap built for all systems and configs.
Aliked + LightGlue is now natively supported.
Note: cuDSS only adds 20% mapping speed and is unstable in some systems.
github.com/lyehe/build_gp…
A new major release of #pySLAM is out ✨
This update pushes the framework forward on:
• Runtime performance →new C++ core
• #SemanticSLAM tightly coupled with volumetric mapping
• End-to-end multi-view 3D scene inference
🧵 Thread below
🔗github.com/luigifreda/pys…
15K Followers 453 FollowingMember of Technical Staff @ DeepSeek, Harness Team
I'd love to connect with members of international frontier LLM labs! DM is open. Opinions are my own.
594K Followers 559 FollowingFounder of the world’s most read daily AI newsletter @therundownai. Sharing the latest developments in the world of artificial intelligence.
1.5M Followers 176 FollowingNobel Laureate. Co-Founder & CEO @GoogleDeepMind - working on AGI. Solving disease @IsomorphicLabs. Trying to understand the fundamental nature of reality.
20K Followers 4 FollowingTweeting interesting papers submitted at https://t.co/rXX8x0HzXV.
Submit your own at https://t.co/QhbJKXBd4Q, and link models/datasets/demos to it!
17K Followers 19 FollowingBuilding a new class of safer, more capable AI systems we call Humanist Superintelligence: AI that is always aligned, controllable, and in service of humanity.
9K Followers 353 FollowingOpenBMB (Open Lab for Big Model Base) aims to build foundation models and systems towards AGI.
Connect with us: https://t.co/N9pevTnoOa
4K Followers 55 FollowingCoze,an AI for For workplace
Discord: https://t.co/37EX8jMHVK
Telegram: https://t.co/jkh53f8uZL
Youtube: https://t.co/Ewy94TcyWt
26K Followers 26 FollowingWe advance the development of ASI and foster open source collaboration towards a smarter future.
Discord: https://t.co/BtsFsAUsvT
11K Followers 164 Following🚀Bringing China's AI & tech trends, voices and perspectives to the global stage.
⚡️Powered by 知乎/https://t.co/OkIemRZdcj, China's leading knowledge community.
810 Followers 85 FollowingDesign & AI Technologies - I design brands and build the AI systems to power them.
Building https://t.co/xeHiCDxiIK (Single Video to 3D Animation)