newsReddit r/MachineLearningTrust 72 · CommunityPublished 1mo agoLive · 1mo ago
Showcase: geolocating a dashcam video without GPS, only from the footage [P]
Sharing a project I have been working on called Third Eye. It does visual geolocation. Given a video, it figures out where it was filmed using only the image content, and draws the route on a map. Pipeline in short: per frame place recognition against a street imagery index a trajectory search that stitches the frames into one coherent path a geometric verification step to catch false matches per frame confiden
Why these links exist
Every edge carries a method, confidence, and the source snippet that justified it — so bad links are debuggable.
- LinkedLinked via unknownRayPE: Ray-Space Positional Encoding for 3D-Aware Video Generation →
- LinkedLinked via unknownMoPe: Motion Permanence for Robust Monocular Gaussian Mapping in Dynamic Environments →
- LinkedLinked via unknown3D Scene-Adaptive Trajectory-Controllable Human Image Animation with Camera Movement →
- PossiblePossibly related (embedding) · 46%GeoMix: Descriptor-Free Visual Localization via Global Context and Multi-Detector Training →
- PossiblePossibly related (embedding) · 48%InFlux++: Real and Synthetic Data for Estimating Dynamic Camera Intrinsics →
- PossiblePossibly related (embedding) · 46%SVI360: Spherical Video Interpolation →
- PossiblePossibly related (embedding) · 47%Self-Healing Visual Recovery for Autonomous Ground Vehicles Using Camera-Only Visual Odometry →
Covers (incoming)
paperRayPE: Ray-Space Positional Encoding for 3D-Aware Video GenerationpaperMoPe: Motion Permanence for Robust Monocular Gaussian Mapping in Dynamic Environmentspaper3D Scene-Adaptive Trajectory-Controllable Human Image Animation with Camera MovementpaperBeyond 2D Matching: A Unified Single-Stage Framework for Geometry-Aware Cross-View Object Geo-LocalizationpaperGeoMix: Descriptor-Free Visual Localization via Global Context and Multi-Detector TrainingpaperInFlux++: Real and Synthetic Data for Estimating Dynamic Camera IntrinsicspaperSVI360: Spherical Video InterpolationpaperSelf-Healing Visual Recovery for Autonomous Ground Vehicles Using Camera-Only Visual OdometrypaperBreaking Déjà Vu: Independent Auditing of Visual Place Recognition through Vision-Language ReasoningpaperToward Semantic Communication for Real-time Mobile 3D ReconstructionpaperRIM: A Retrieval-In-Matching Framework for Cross-Domain Global Visual Localization of UAVspaperTraVEL: Trajectory-Guided Video Embedding Learning for Driving-Video Retrieval
Related across the graph
paperTraVEL: Trajectory-Guided Video Embedding Learning for Driving-Video RetrievalpaperToward Semantic Communication for Real-time Mobile 3D Reconstructionpaper3D Scene-Adaptive Trajectory-Controllable Human Image Animation with Camera MovementpaperRayPE: Ray-Space Positional Encoding for 3D-Aware Video GenerationpaperMoPe: Motion Permanence for Robust Monocular Gaussian Mapping in Dynamic EnvironmentspaperBeyond 2D Matching: A Unified Single-Stage Framework for Geometry-Aware Cross-View Object Geo-LocalizationpaperGeoMix: Descriptor-Free Visual Localization via Global Context and Multi-Detector TrainingpaperBreaking Déjà Vu: Independent Auditing of Visual Place Recognition through Vision-Language ReasoningpaperSelf-Healing Visual Recovery for Autonomous Ground Vehicles Using Camera-Only Visual OdometrypaperInFlux++: Real and Synthetic Data for Estimating Dynamic Camera IntrinsicspaperSVI360: Spherical Video InterpolationpaperRIM: A Retrieval-In-Matching Framework for Cross-Domain Global Visual Localization of UAVs
