[2006.11275] Center-based 3D Object Detection and Tracking
arXivLabs is a framework that allows collaborators to develop and share new arXiv features directly on our website. Both individuals and organizations that work with arXivLabs have embraced and accepted our values of openness, community, excellence, and user data privacy. arXiv is committed to these values and only works with partners that adhere to them. Have an idea for a project that will add value for arXiv's community? Learn more about arXivLabs. arXiv Operational Status Get status notifications via email or slack
[2006.11275] Center-based 3D Object Detection and Tracking Skip to main content arXiv is now an independent nonprofit! Learn more × Search arXiv Press Enter to search · Advanced search --> Computer Science > Computer Vision and Pattern Recognition arXiv:2006.11275 (cs) [Submitted on 19 Jun 2020 ( v1 ), last revised 6 Jan 2021 (this version, v2)] Title: Center-based 3D Object Detection and Tracking Authors: Tianwei Yin , Xingyi Zhou , Philipp Krähenbühl View a PDF of the paper titled Center-based 3D Object Detection and Tracking, by Tianwei Yin and 2 other authors View PDF Abstract
Explore this link on the map →related reading
- [1711.06396] VoxelNet: End-to-End Learning for Point Cloud Based 3D Object Detectionarxiv.org
- ReferIt3D: Neural Listeners for Fine-Grained 3D Object Identification in Real-World Scenesecva.net
- GitHub - xingyizhou/ExtremeNet: Bottom-up Object Detection by Grouping Extreme and Center Points · GitHubgithub.com
- ReferIt3D Benchmarksreferit3d.github.io
- Multi-View Transformer for 3D Visual Groundingarxiv.org
- Reference for ultralytics/trackers/bot_sort.py | Ultralytics Docsdocs.ultralytics.com
- 3D-LFM: Lifting Foundation Model3dlfm.github.io
- [2310.08586] PonderV2: Pave the Way for 3D Foundation Model with A Universal Pre-training Paradigmarxiv.org
- [2008.05711] Lift, Splat, Shoot: Encoding Images From Arbitrary Camera Rigs by Implicitly Unprojecting to 3Darxiv.org
- PLA: Language-Driven Open-Vocabulary 3D Scene Understandingarxiv.org
- Lift, Splat, Shoot: Encoding Images from Arbitrary Camera Rigs by Implicitly Unprojecting to 3Dresearch.nvidia.com
- The flavor of the bitter lesson for computer vision - Vincent Sitzmannvincentsitzmann.com