PaperLens
紙
Students
Professional
JA
EN
◐
Sign in with Google
Sign in
Read
Home
Close reading
New
Textbook
Go deeper
Learn
Lab
Landscape
Contributors
Glossary
You
Search
All-access
My Page
#camera-geometry
1 articles
01
2026-08-12
·
VLMs & Multimodal
·
★ MEMBER
·
PAPER
·
8 min read
BEV Representations From Scratch — Fusing Multiple Cameras Into One Top-Down Map
How a self-driving car turns six camera feeds into a single top-down map. Starting from perspective projection, we build up to the two big design philosophies: LSS, which pushes features into 3D via a predicted depth distribution, and Transformer-style methods like BEVFormer that pull information with BEV queries.