Pith. sign in

Paper Citation Record · LEDGER

MVL-Loc: Leveraging Vision-Language Model for Generalizable Multi-Scene Camera Relocalization

As of 10 August 2026, this Paper Citation Record lists 36 of 36 outbound references and 0 inbound Pith citation observations for arXiv:2507.04509.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2507.04509 v1

Coverage vector

measured 36 of 36 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-06T19:51:53.074321Z

measured 36 of 36 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

36 of 36 outbound references displayed

  • verified exact0
  • verified fuzzy31
  • unresolved5
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 638fde49-e415-43b4-bea2-c236ad418035 · outbound

This paper cites Random sample consensus: A paradigm for model fitting with applications to image analysis and automated cartography,.

MVL-Loc: Leveraging Vision-Language Model for Generalizable Multi-Scene Camera Relocalization Random sample consensus: A paradigm for model fitting with applications to image analysis and automated cartography,

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:51:58.329827Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T19:51:50.260692Z digest=sha256:257619144ee36585255c0e8f265f75d2e84091fd79d359c0b5294c8f90421b29

Observation 362dd15c-f339-4545-95e0-e0491b5f214d · outbound

This paper cites Posenet: A convolutional network for real-time 6-dof camera relocalization,.

MVL-Loc: Leveraging Vision-Language Model for Generalizable Multi-Scene Camera Relocalization Posenet: A convolutional network for real-time 6-dof camera relocalization,

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:51:58.175786Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T19:51:50.355729Z digest=sha256:8d2318cee52750c8c63a5617b835fc440e388b9b7760cebbf436407bc8c4928a

Observation 6e56b12d-7ddd-47d2-8c99-9342ec23b794 · outbound

This paper cites Image-based localization using lstms for structured feature correlation,.

MVL-Loc: Leveraging Vision-Language Model for Generalizable Multi-Scene Camera Relocalization Image-based localization using lstms for structured feature correlation,

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:51:58.050472Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T19:51:50.425621Z digest=sha256:090dd4be52ca7879eceada9aa53ab38c73e72f00c69c3de087c53ce52c101d0c

Observation d23f89fb-7306-462d-b63d-0850a8f6cd66 · outbound

This paper cites Image-based localization using hourglass networks,.

MVL-Loc: Leveraging Vision-Language Model for Generalizable Multi-Scene Camera Relocalization Image-based localization using hourglass networks,

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:51:57.962733Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T19:51:50.505453Z digest=sha256:16125e1481a0069b5be247eace397fc4ff0f9fd5df4dfc368841ada0147cd065

Observation 96a5534c-f03e-4db4-9c85-9049bf354677 · outbound

This paper cites Modelling uncertainty in deep learning for camera relocalization,.

MVL-Loc: Leveraging Vision-Language Model for Generalizable Multi-Scene Camera Relocalization Modelling uncertainty in deep learning for camera relocalization,

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:51:57.814913Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T19:51:50.541883Z digest=sha256:1c090f50dd2231f692e96d2712341bf77c6cbaa79f62ccf2a5183d7ec56e8b09

Observation 7d8c9377-86e9-4e77-b200-acd37f6af818 · outbound

This paper cites Geometric loss functions for camera pose regression with deep learning,.

MVL-Loc: Leveraging Vision-Language Model for Generalizable Multi-Scene Camera Relocalization Geometric loss functions for camera pose regression with deep learning,

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:51:57.672985Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T19:51:50.649083Z digest=sha256:65de78f012c0d06be8088418190cd825e42b41d8bc238041885854de639f5ba4

Observation f5555b2d-9c3e-4b87-825d-3a101907eee6 · outbound

This paper cites Vidloc: A deep spatio-temporal model for 6-dof video-clip relocalization,.

MVL-Loc: Leveraging Vision-Language Model for Generalizable Multi-Scene Camera Relocalization Vidloc: A deep spatio-temporal model for 6-dof video-clip relocalization,

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:51:57.529663Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T19:51:50.708062Z digest=sha256:9c054a814010fa615619113107e9dd2b45dd8466694d9a020f758392b4fe1bdc

Observation c35f19f2-66a0-4abc-b7f7-82da0b72176c · outbound

This paper cites Atloc: Attention guided camera localization,.

MVL-Loc: Leveraging Vision-Language Model for Generalizable Multi-Scene Camera Relocalization Atloc: Attention guided camera localization,

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:51:57.416565Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T19:51:50.797015Z digest=sha256:710a1f575b2031ab53673ce7da7fc9f401f02a81e2b59a73e558673e79a279d7

Observation 91c298e2-af3a-43ba-8ed9-cfd3119541c6 · outbound

This paper cites Effloc: Lightweight vision transformer for effi- cient 6-dof camera relocalization,.

MVL-Loc: Leveraging Vision-Language Model for Generalizable Multi-Scene Camera Relocalization Effloc: Lightweight vision transformer for effi- cient 6-dof camera relocalization,

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:51:57.300384Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T19:51:50.878081Z digest=sha256:9e06b73b247c9e8e5b6b2256c30b9e0dd523d1c5dca91c03076f78fdcb95eb13

Observation 11c3bad6-0b8d-4f07-b6de-e5207a512cf4 · outbound

This paper cites NavGPT: Explicit Reasoning in Vision-and-Language Navigation with Large Language Models.

MVL-Loc: Leveraging Vision-Language Model for Generalizable Multi-Scene Camera Relocalization NavGPT: Explicit Reasoning in Vision-and-Language Navigation with Large Language Models

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-06T19:51:50.963143Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:51:50.963143Z digest=sha256:f5def470a30176f17a60ac83fb6888d53212e74edd0192afd5dc5e68c1014bf2

Observation 8a0113dc-d0f3-4fa9-97b9-53026008ccf0 · outbound

This paper cites RT-2: Vision-Language-Action Models Transfer Web Knowledge to Robotic Control.

MVL-Loc: Leveraging Vision-Language Model for Generalizable Multi-Scene Camera Relocalization RT-2: Vision-Language-Action Models Transfer Web Knowledge to Robotic Control

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-06T19:51:51.065164Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:51:51.065164Z digest=sha256:60e6a384ac6f708a37b2373439d6688cffec9fa940d63a738c18567deaeb50ec

Observation ef2c3e76-431d-4d3d-a9c4-372d8b0ea5f7 · outbound

This paper cites Extending absolute pose regression to multiple scenes,.

MVL-Loc: Leveraging Vision-Language Model for Generalizable Multi-Scene Camera Relocalization Extending absolute pose regression to multiple scenes,

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:51:57.175742Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T19:51:51.145193Z digest=sha256:602162265e2ef85061d912bf9fe8ab9d2c5749af0b0000d314d8c291cbfcea1d

Observation 3c40f7ac-7a35-4f5d-ba29-539d8f14ddcd · outbound

This paper cites Coarse-to-fine multi-scene pose regression with trans- formers,.

MVL-Loc: Leveraging Vision-Language Model for Generalizable Multi-Scene Camera Relocalization Coarse-to-fine multi-scene pose regression with trans- formers,

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:51:57.053636Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T19:51:51.218239Z digest=sha256:cbacbd3b8b94389d1377c73bef54ed4b09bb3c48d6da070328271a3377be7573

Observation 6df5821e-0bac-4bdf-86b9-8ac1b311214a · outbound

This paper cites City-scale landmark identification on mobile devices,.

MVL-Loc: Leveraging Vision-Language Model for Generalizable Multi-Scene Camera Relocalization City-scale landmark identification on mobile devices,

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:51:56.892027Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T19:51:51.307081Z digest=sha256:c922d8e655b6cbc82338d595e2f64cdda1af4a011758e19cdf53005c3eada984

Observation c13af43b-2b0c-4343-b543-574fb85d132c · outbound

This paper cites Imagdressing-v1: Cus- tomizable virtual dressing,.

MVL-Loc: Leveraging Vision-Language Model for Generalizable Multi-Scene Camera Relocalization Imagdressing-v1: Cus- tomizable virtual dressing,

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:51:56.661234Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T19:51:51.394012Z digest=sha256:bfdc373ad0aaa24ddf5902dfb3568cb79fe34239452f7804e73e9ea020f6ca24

Observation 3243608f-e3ab-4310-9934-8b36b0961e8b · outbound

This paper cites Imagpose: A unified conditional framework for pose-guided person generation,.

MVL-Loc: Leveraging Vision-Language Model for Generalizable Multi-Scene Camera Relocalization Imagpose: A unified conditional framework for pose-guided person generation,

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-06T19:51:51.465976Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:51:51.465976Z digest=sha256:fae9187e69ba767664bfb6d80ff6cc23874fd3fdd09e5b3e4d4c089d0e035bcd

Observation 28313878-7355-4dd8-88e7-2d146d3d4dd9 · outbound

This paper cites Dsac-differentiable ransac for camera localization,.

MVL-Loc: Leveraging Vision-Language Model for Generalizable Multi-Scene Camera Relocalization Dsac-differentiable ransac for camera localization,

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:51:56.366071Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T19:51:51.567426Z digest=sha256:dfe1cd8bfec2cd0c2c4d3a88b79b9dbedb91219c71ec7eac31166d3687f5faef

Observation eaf78365-3f39-4339-94d2-fb681fe3373e · outbound

This paper cites Hybrid scene compression for visual localization,.

MVL-Loc: Leveraging Vision-Language Model for Generalizable Multi-Scene Camera Relocalization Hybrid scene compression for visual localization,

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:51:56.099167Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T19:51:51.604086Z digest=sha256:f7cf902e6561eaca0da28179078dcda89af2d9e2e87ee3355d2b3c05173f1083

Observation c0b92286-dabf-42d7-a745-590ecbcb0bc2 · outbound

This paper cites Map-free visual relocalization: Metric pose relative to a single image,.

MVL-Loc: Leveraging Vision-Language Model for Generalizable Multi-Scene Camera Relocalization Map-free visual relocalization: Metric pose relative to a single image,

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:51:55.912782Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T19:51:51.641587Z digest=sha256:e785e1f56d2cb9cc59f06fb5402ba871586f1dae6b075747210ee82eaa9a79da

Observation cec9be87-48db-4302-8628-d1ced13e8a31 · outbound

This paper cites Learning multi-scene absolute pose regression with transformers,.

MVL-Loc: Leveraging Vision-Language Model for Generalizable Multi-Scene Camera Relocalization Learning multi-scene absolute pose regression with transformers,

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:51:55.740171Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T19:51:51.742970Z digest=sha256:036b79069a14c2a4e834ea7a5bcfb200c4722f539bf4f1a75a70c745cde57d7e

Observation cc0e3ccf-ea5f-405f-b71a-94df2dabad27 · outbound

This paper cites Geometry-aware learning of maps for camera localization,.

MVL-Loc: Leveraging Vision-Language Model for Generalizable Multi-Scene Camera Relocalization Geometry-aware learning of maps for camera localization,

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:51:55.618580Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T19:51:51.798687Z digest=sha256:5f3fe25d2a4bf844b837134a5d524d38b35de286ed3b1d0a49c930990d3c5b52

Observation c44a2cfe-b1cc-4474-941d-dcaf2fd101b9 · outbound

This paper cites Fusionloc: Camera-2d lidar fusion using multi-head self-attention for end-to-end serving robot relocalization,.

MVL-Loc: Leveraging Vision-Language Model for Generalizable Multi-Scene Camera Relocalization Fusionloc: Camera-2d lidar fusion using multi-head self-attention for end-to-end serving robot relocalization,

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:51:55.452563Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T19:51:51.894716Z digest=sha256:e7f2b1a6fa68461548b375c2c6a233b4549d2e8119315682ccaab6ff0cdbe906

Observation 75f786dc-6c23-417d-8104-ed146e475ff7 · outbound

This paper cites Learning transferable visual models from natural language supervision,.

MVL-Loc: Leveraging Vision-Language Model for Generalizable Multi-Scene Camera Relocalization Learning transferable visual models from natural language supervision,

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:51:55.249451Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T19:51:51.949874Z digest=sha256:9ab875ee89cd334a7d0e872951ba8540d011536f8038f10fef168e3c548ee3ad

Observation 7b745bcb-5c88-475d-b6fc-c15536b35469 · outbound

This paper cites CLIP-Adapter: Better Vision-Language Models with Feature Adapters.

MVL-Loc: Leveraging Vision-Language Model for Generalizable Multi-Scene Camera Relocalization CLIP-Adapter: Better Vision-Language Models with Feature Adapters

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-06T19:51:52.021631Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:51:52.021631Z digest=sha256:07548571e0295d6ed168c092c02f5ccb5738acef9a967d98cb5c25990212eea8

Observation e814071b-bf23-4026-8066-692a376fc4c7 · outbound

This paper cites Envedit: Environment editing for vision-and-language naviga- tion,.

MVL-Loc: Leveraging Vision-Language Model for Generalizable Multi-Scene Camera Relocalization Envedit: Environment editing for vision-and-language naviga- tion,

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:51:55.073543Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T19:51:52.093706Z digest=sha256:aff1f359040c79b8f20f73b75b7b2d1026caebfff3db592731d36769b32b8084

Observation eb68e5e1-1336-4a25-be70-1f42839c86d7 · outbound

This paper cites Learning to generate scene graph from natural language supervision,.

MVL-Loc: Leveraging Vision-Language Model for Generalizable Multi-Scene Camera Relocalization Learning to generate scene graph from natural language supervision,

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:51:54.966464Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T19:51:52.204735Z digest=sha256:a1eb31f91bffd2998ed0c7e992ec77aa0718cd8f99fda9aeb97b73dc4924d49d

Observation c666e258-ce83-4cdc-87b0-d6f1eb17bad3 · outbound

This paper cites Denseclip: Language-guided dense prediction with context-aware prompting,.

MVL-Loc: Leveraging Vision-Language Model for Generalizable Multi-Scene Camera Relocalization Denseclip: Language-guided dense prediction with context-aware prompting,

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:51:54.855050Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T19:51:52.282022Z digest=sha256:8b7d7df3641691db721a1d10209e1f40d52a2099b2bf5018f76a193cb2372925

Observation 9afee136-979f-4461-ad99-dda00c0805d6 · outbound

This paper cites Fm-loc: Using foundation models for improved vision-based localization,.

MVL-Loc: Leveraging Vision-Language Model for Generalizable Multi-Scene Camera Relocalization Fm-loc: Using foundation models for improved vision-based localization,

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:51:54.708146Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T19:51:52.397707Z digest=sha256:a778be186a1745143b11e5964ff82af27e681bc95acc0746dbc9b808eb7b4c26

Observation b37d3e6a-e5fd-4118-b89b-61be37dd36e0 · outbound

This paper cites Clip-loc: Multi-modal landmark association for global localization in object-based maps,.

MVL-Loc: Leveraging Vision-Language Model for Generalizable Multi-Scene Camera Relocalization Clip-loc: Multi-modal landmark association for global localization in object-based maps,

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:51:54.488600Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T19:51:52.488586Z digest=sha256:a38ecadcebd0edab73b2a768ae249828c5b78eea0a2372164e15970653b0ed8a

Observation f1622f84-0953-49c4-8d5a-4610bba884a6 · outbound

This paper cites Geollm: Extracting geospatial knowledge from large language models,.

MVL-Loc: Leveraging Vision-Language Model for Generalizable Multi-Scene Camera Relocalization Geollm: Extracting geospatial knowledge from large language models,

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:51:54.312406Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T19:51:52.561556Z digest=sha256:8d7ace2da2a822b30fcc5926149c1c97848d5b649ed4d332557553765b18f51f

Observation 40438513-e33f-406f-9f44-371719b0664f · outbound

This paper cites Georeasoner: Geo-localization with reasoning in street views using a large vision-language model,.

MVL-Loc: Leveraging Vision-Language Model for Generalizable Multi-Scene Camera Relocalization Georeasoner: Geo-localization with reasoning in street views using a large vision-language model,

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:51:54.112020Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T19:51:52.660280Z digest=sha256:ff7453ec8681c8d175ff2e69f2ebc5e81269cd8528ea909a566e05204cded73d

Observation 914a639e-3631-4c88-8f43-7a73f31c2c34 · outbound

This paper cites Advancing Pose-Guided Image Synthesis with Progressive Conditional Diffusion Models.

MVL-Loc: Leveraging Vision-Language Model for Generalizable Multi-Scene Camera Relocalization Advancing Pose-Guided Image Synthesis with Progressive Conditional Diffusion Models

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-06T19:51:52.756841Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:51:52.756841Z digest=sha256:41110026f2ca322420020dcf32fc112fcd2b27182a0d4605d5ab1a1113ba1ef0

Observation 64f9f7c1-362b-477c-a6a6-2db52a0c705f · outbound

This paper cites Boosting consistency in story visualization with rich-contextual conditional diffusion models,.

MVL-Loc: Leveraging Vision-Language Model for Generalizable Multi-Scene Camera Relocalization Boosting consistency in story visualization with rich-contextual conditional diffusion models,

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:51:53.933474Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T19:51:52.847978Z digest=sha256:8557d8d903a60c75effb71c8c754a3ffbf571dd8940ba0d78f210d589aa5e0e4

Observation 2e3092c5-4a38-4a7d-a651-017588757960 · outbound

This paper cites 7-scenes dataset,.

MVL-Loc: Leveraging Vision-Language Model for Generalizable Multi-Scene Camera Relocalization 7-scenes dataset,

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:51:53.706874Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T19:51:52.928410Z digest=sha256:a589cde74cf8d6340ad41974a93298ea28f360814445f810b5765bd713f618c8

Observation cbb96d11-fce4-4cb5-9e62-0dc5cccb8b9c · outbound

This paper cites Do we really need scene-specific pose encoders?.

MVL-Loc: Leveraging Vision-Language Model for Generalizable Multi-Scene Camera Relocalization Do we really need scene-specific pose encoders?

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:51:53.521611Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T19:51:53.005097Z digest=sha256:765f5047f06dafc81167ce71ca0e6810d1cf5f0bd34776fb19c991364b39a42a

Observation d4c93214-bca1-4038-99db-956f9694ab88 · outbound

This paper cites Image-based localization using hourglass networks,.

MVL-Loc: Leveraging Vision-Language Model for Generalizable Multi-Scene Camera Relocalization Image-based localization using hourglass networks,

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:51:53.328502Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T19:51:53.074321Z digest=sha256:54d2cb223eac3d1369c31326d59faceddecdf9bb3f77feb53e8f925ccb94af45

Pith citing papers

No inbound Pith citation observations are available.