Pith. sign in

Paper Citation Record · LEDGER

MVL-Loc: Leveraging Vision-Language Model for Generalizable Multi-Scene Camera Relocalization

As of 17 August 2026, this Paper Citation Record lists 36 of 36 outbound references and 0 inbound Pith citation observations for arXiv:2507.04509.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2507.04509 v1

Coverage vector

measured 36 of 36 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-06T19:51:53.074321Z

measured 36 of 36 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-16T06:30:59.297886+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

36 of 36 outbound references displayed

  • verified exact0
  • verified fuzzy31
  • unresolved5
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 638fde49-e415-43b4-bea2-c236ad418035 · outbound

This paper cites Random sample consensus: A paradigm for model fitting with applications to image analysis and automated cartography,.

MVL-Loc: Leveraging Vision-Language Model for Generalizable Multi-Scene Camera Relocalization Random sample consensus: A paradigm for model fitting with applications to image analysis and automated cartography,

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:51:58.329827Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-06T19:51:50.260692Z digest=sha256:d76f811a2552b367412166a398a94699eeb7be6e03cce03834f0df12fd97fa5d

Observation 362dd15c-f339-4545-95e0-e0491b5f214d · outbound

This paper cites Posenet: A convolutional network for real-time 6-dof camera relocalization,.

MVL-Loc: Leveraging Vision-Language Model for Generalizable Multi-Scene Camera Relocalization Posenet: A convolutional network for real-time 6-dof camera relocalization,

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:51:58.175786Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-06T19:51:50.355729Z digest=sha256:6a2269e0823c7400fae0e09d1cb2f0bb2d8d2c1d1fa48a383647ba6937f391fe

Observation 6e56b12d-7ddd-47d2-8c99-9342ec23b794 · outbound

This paper cites Image-based localization using lstms for structured feature correlation,.

MVL-Loc: Leveraging Vision-Language Model for Generalizable Multi-Scene Camera Relocalization Image-based localization using lstms for structured feature correlation,

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:51:58.050472Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-06T19:51:50.425621Z digest=sha256:f477687b180ab28bb7470ed02e238f50bca62cb24862784871ebb46fc806b47b

Observation d23f89fb-7306-462d-b63d-0850a8f6cd66 · outbound

This paper cites Image-based localization using hourglass networks,.

MVL-Loc: Leveraging Vision-Language Model for Generalizable Multi-Scene Camera Relocalization Image-based localization using hourglass networks,

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:51:57.962733Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-06T19:51:50.505453Z digest=sha256:773aeb1bedc2de115ba172bcbe5018dc3b7369bd6c29459bc2ca81675475545f

Observation 96a5534c-f03e-4db4-9c85-9049bf354677 · outbound

This paper cites Modelling uncertainty in deep learning for camera relocalization,.

MVL-Loc: Leveraging Vision-Language Model for Generalizable Multi-Scene Camera Relocalization Modelling uncertainty in deep learning for camera relocalization,

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:51:57.814913Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-06T19:51:50.541883Z digest=sha256:b2bdafe892d13bd3ebaf3c86320e765e328cc20e8ed00cee1732d666e468df72

Observation 7d8c9377-86e9-4e77-b200-acd37f6af818 · outbound

This paper cites Geometric loss functions for camera pose regression with deep learning,.

MVL-Loc: Leveraging Vision-Language Model for Generalizable Multi-Scene Camera Relocalization Geometric loss functions for camera pose regression with deep learning,

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:51:57.672985Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-06T19:51:50.649083Z digest=sha256:6487101351a3f84cbcdf367fb9e0b22a2334eb67dcdedac8f764caace964856a

Observation f5555b2d-9c3e-4b87-825d-3a101907eee6 · outbound

This paper cites Vidloc: A deep spatio-temporal model for 6-dof video-clip relocalization,.

MVL-Loc: Leveraging Vision-Language Model for Generalizable Multi-Scene Camera Relocalization Vidloc: A deep spatio-temporal model for 6-dof video-clip relocalization,

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:51:57.529663Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-06T19:51:50.708062Z digest=sha256:2176bdee4581f8be20f9112b8edb8f1ce6a9a4c67f57001f73e4a84ab31f0e4e

Observation c35f19f2-66a0-4abc-b7f7-82da0b72176c · outbound

This paper cites Atloc: Attention guided camera localization,.

MVL-Loc: Leveraging Vision-Language Model for Generalizable Multi-Scene Camera Relocalization Atloc: Attention guided camera localization,

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:51:57.416565Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-06T19:51:50.797015Z digest=sha256:8a6c14169a3390a343b3a39cfacabd415cabea21f136264f58a715a25d933b3b

Observation 91c298e2-af3a-43ba-8ed9-cfd3119541c6 · outbound

This paper cites Effloc: Lightweight vision transformer for effi- cient 6-dof camera relocalization,.

MVL-Loc: Leveraging Vision-Language Model for Generalizable Multi-Scene Camera Relocalization Effloc: Lightweight vision transformer for effi- cient 6-dof camera relocalization,

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:51:57.300384Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-06T19:51:50.878081Z digest=sha256:b1581cb35676e21a871fa7b58b477bdf32a8d4fbc472c19578738aa069225d0d

Observation 11c3bad6-0b8d-4f07-b6de-e5207a512cf4 · outbound

This paper cites NavGPT: Explicit Reasoning in Vision-and-Language Navigation with Large Language Models.

MVL-Loc: Leveraging Vision-Language Model for Generalizable Multi-Scene Camera Relocalization NavGPT: Explicit Reasoning in Vision-and-Language Navigation with Large Language Models

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-06T19:51:50.963143Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:51:50.963143Z digest=sha256:635c0f0d4e7cafd459060ea016495df949ced90841aff27f5f0b79ecf47b6e74

Observation 8a0113dc-d0f3-4fa9-97b9-53026008ccf0 · outbound

This paper cites RT-2: Vision-Language-Action Models Transfer Web Knowledge to Robotic Control.

MVL-Loc: Leveraging Vision-Language Model for Generalizable Multi-Scene Camera Relocalization RT-2: Vision-Language-Action Models Transfer Web Knowledge to Robotic Control

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-06T19:51:51.065164Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:51:51.065164Z digest=sha256:0e40aaec8c43f45f0ae6a4bd13d6fa8f844da019fd81b30f9627d833035f7163

Observation ef2c3e76-431d-4d3d-a9c4-372d8b0ea5f7 · outbound

This paper cites Extending absolute pose regression to multiple scenes,.

MVL-Loc: Leveraging Vision-Language Model for Generalizable Multi-Scene Camera Relocalization Extending absolute pose regression to multiple scenes,

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:51:57.175742Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-06T19:51:51.145193Z digest=sha256:8e8383af2e4864a088d135006610f6ebe90b0a2d7976e48ea169fd72f9bc2add

Observation 3c40f7ac-7a35-4f5d-ba29-539d8f14ddcd · outbound

This paper cites Coarse-to-fine multi-scene pose regression with trans- formers,.

MVL-Loc: Leveraging Vision-Language Model for Generalizable Multi-Scene Camera Relocalization Coarse-to-fine multi-scene pose regression with trans- formers,

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:51:57.053636Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-06T19:51:51.218239Z digest=sha256:d879946867bf35b8769918b2734476c4607a33f495843d84a1effe21cf159923

Observation 6df5821e-0bac-4bdf-86b9-8ac1b311214a · outbound

This paper cites City-scale landmark identification on mobile devices,.

MVL-Loc: Leveraging Vision-Language Model for Generalizable Multi-Scene Camera Relocalization City-scale landmark identification on mobile devices,

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:51:56.892027Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-06T19:51:51.307081Z digest=sha256:e40e2d6a2e22cce77196711b778f98cc45de5b86485173b8a40acabc179f2a3a

Observation c13af43b-2b0c-4343-b543-574fb85d132c · outbound

This paper cites Imagdressing-v1: Cus- tomizable virtual dressing,.

MVL-Loc: Leveraging Vision-Language Model for Generalizable Multi-Scene Camera Relocalization Imagdressing-v1: Cus- tomizable virtual dressing,

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:51:56.661234Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-06T19:51:51.394012Z digest=sha256:9a2f0f803fdd76436d42d2c301ebee9ccd748ee0e34e121dfe039e5b7438ce66

Observation 3243608f-e3ab-4310-9934-8b36b0961e8b · outbound

This paper cites Imagpose: A unified conditional framework for pose-guided person generation,.

MVL-Loc: Leveraging Vision-Language Model for Generalizable Multi-Scene Camera Relocalization Imagpose: A unified conditional framework for pose-guided person generation,

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-06T19:51:51.465976Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:51:51.465976Z digest=sha256:0feaecc784e94aa17d3989bbc1a91dbd0d290c31463625948f6573bfb5281665

Observation 28313878-7355-4dd8-88e7-2d146d3d4dd9 · outbound

This paper cites Dsac-differentiable ransac for camera localization,.

MVL-Loc: Leveraging Vision-Language Model for Generalizable Multi-Scene Camera Relocalization Dsac-differentiable ransac for camera localization,

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:51:56.366071Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-06T19:51:51.567426Z digest=sha256:cff6ae2c2285be64b60d378a06f72c45c3f395c59d9229ee5a7a3d54a54a8d86

Observation eaf78365-3f39-4339-94d2-fb681fe3373e · outbound

This paper cites Hybrid scene compression for visual localization,.

MVL-Loc: Leveraging Vision-Language Model for Generalizable Multi-Scene Camera Relocalization Hybrid scene compression for visual localization,

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:51:56.099167Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-06T19:51:51.604086Z digest=sha256:c5c2af342d6b3d6eef4e329c7ab48eca1561b2e6ce488debcae613f9b53bcdb2

Observation c0b92286-dabf-42d7-a745-590ecbcb0bc2 · outbound

This paper cites Map-free visual relocalization: Metric pose relative to a single image,.

MVL-Loc: Leveraging Vision-Language Model for Generalizable Multi-Scene Camera Relocalization Map-free visual relocalization: Metric pose relative to a single image,

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:51:55.912782Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-06T19:51:51.641587Z digest=sha256:d63b203acb1fe7ea0e54e54ff4803617d2a4a35d1e90c0efb47c0725b9bc183c

Observation cec9be87-48db-4302-8628-d1ced13e8a31 · outbound

This paper cites Learning multi-scene absolute pose regression with transformers,.

MVL-Loc: Leveraging Vision-Language Model for Generalizable Multi-Scene Camera Relocalization Learning multi-scene absolute pose regression with transformers,

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:51:55.740171Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-06T19:51:51.742970Z digest=sha256:8aacc499fece402aaf95e162126ee0111a6fb99ae2495af9efe65e35e945274c

Observation cc0e3ccf-ea5f-405f-b71a-94df2dabad27 · outbound

This paper cites Geometry-aware learning of maps for camera localization,.

MVL-Loc: Leveraging Vision-Language Model for Generalizable Multi-Scene Camera Relocalization Geometry-aware learning of maps for camera localization,

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:51:55.618580Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-06T19:51:51.798687Z digest=sha256:4d4d13bbfc52260148e73f159ad19daf1ffebb3a9fccb383628a8e5f7c3d1a09

Observation c44a2cfe-b1cc-4474-941d-dcaf2fd101b9 · outbound

This paper cites Fusionloc: Camera-2d lidar fusion using multi-head self-attention for end-to-end serving robot relocalization,.

MVL-Loc: Leveraging Vision-Language Model for Generalizable Multi-Scene Camera Relocalization Fusionloc: Camera-2d lidar fusion using multi-head self-attention for end-to-end serving robot relocalization,

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:51:55.452563Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-06T19:51:51.894716Z digest=sha256:f9f49fee3a7b523002c3c17f84aba42c4490c2580a58958826109ec44a909f87

Observation 75f786dc-6c23-417d-8104-ed146e475ff7 · outbound

This paper cites Learning transferable visual models from natural language supervision,.

MVL-Loc: Leveraging Vision-Language Model for Generalizable Multi-Scene Camera Relocalization Learning transferable visual models from natural language supervision,

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:51:55.249451Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-06T19:51:51.949874Z digest=sha256:c091423f7cbcff7eaad2dabb8a4e9cf2a0e8372f390b18fa4f34b812315691ad

Observation 7b745bcb-5c88-475d-b6fc-c15536b35469 · outbound

This paper cites CLIP-Adapter: Better Vision-Language Models with Feature Adapters.

MVL-Loc: Leveraging Vision-Language Model for Generalizable Multi-Scene Camera Relocalization CLIP-Adapter: Better Vision-Language Models with Feature Adapters

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-06T19:51:52.021631Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:51:52.021631Z digest=sha256:d4f70a64bacbfb9b40467159c504dcb19327cc5a5521134594d3e695738c9c0a

Observation e814071b-bf23-4026-8066-692a376fc4c7 · outbound

This paper cites Envedit: Environment editing for vision-and-language naviga- tion,.

MVL-Loc: Leveraging Vision-Language Model for Generalizable Multi-Scene Camera Relocalization Envedit: Environment editing for vision-and-language naviga- tion,

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:51:55.073543Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-06T19:51:52.093706Z digest=sha256:c21dbac12b53f6347e3976c55a4ec13dd3390dfbccb0d202f357cf545545baed

Observation eb68e5e1-1336-4a25-be70-1f42839c86d7 · outbound

This paper cites Learning to generate scene graph from natural language supervision,.

MVL-Loc: Leveraging Vision-Language Model for Generalizable Multi-Scene Camera Relocalization Learning to generate scene graph from natural language supervision,

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:51:54.966464Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-06T19:51:52.204735Z digest=sha256:ffba236bf681c8d2fba86fbe87a5ece4b6e6f7fc5a0f1a53fdca86ed04fe722a

Observation c666e258-ce83-4cdc-87b0-d6f1eb17bad3 · outbound

This paper cites Denseclip: Language-guided dense prediction with context-aware prompting,.

MVL-Loc: Leveraging Vision-Language Model for Generalizable Multi-Scene Camera Relocalization Denseclip: Language-guided dense prediction with context-aware prompting,

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:51:54.855050Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-06T19:51:52.282022Z digest=sha256:c0ee427e63ec7b7ad13a246bcbb6c9cd56795d8b7585d617c6bae94b32c75863

Observation 9afee136-979f-4461-ad99-dda00c0805d6 · outbound

This paper cites Fm-loc: Using foundation models for improved vision-based localization,.

MVL-Loc: Leveraging Vision-Language Model for Generalizable Multi-Scene Camera Relocalization Fm-loc: Using foundation models for improved vision-based localization,

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:51:54.708146Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-06T19:51:52.397707Z digest=sha256:6185ad31fa0a60fe2d211b061bce137b98e70e14a09559f79e1e4df19436b135

Observation b37d3e6a-e5fd-4118-b89b-61be37dd36e0 · outbound

This paper cites Clip-loc: Multi-modal landmark association for global localization in object-based maps,.

MVL-Loc: Leveraging Vision-Language Model for Generalizable Multi-Scene Camera Relocalization Clip-loc: Multi-modal landmark association for global localization in object-based maps,

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:51:54.488600Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-06T19:51:52.488586Z digest=sha256:719cc7bd0e36abfd059f410f253e667f67ef28b36ffebd9c9089b01109e3390b

Observation f1622f84-0953-49c4-8d5a-4610bba884a6 · outbound

This paper cites Geollm: Extracting geospatial knowledge from large language models,.

MVL-Loc: Leveraging Vision-Language Model for Generalizable Multi-Scene Camera Relocalization Geollm: Extracting geospatial knowledge from large language models,

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:51:54.312406Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-06T19:51:52.561556Z digest=sha256:a04c9314c0e4a1c6ef266756abefa43c9d881961d92065a52a912f5f49e0251b

Observation 40438513-e33f-406f-9f44-371719b0664f · outbound

This paper cites Georeasoner: Geo-localization with reasoning in street views using a large vision-language model,.

MVL-Loc: Leveraging Vision-Language Model for Generalizable Multi-Scene Camera Relocalization Georeasoner: Geo-localization with reasoning in street views using a large vision-language model,

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:51:54.112020Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-06T19:51:52.660280Z digest=sha256:56e0d36f61aa4eb4b50ebfcc16dc92ad32dfc18269e88f12962a79675450c647

Observation 914a639e-3631-4c88-8f43-7a73f31c2c34 · outbound

This paper cites Advancing Pose-Guided Image Synthesis with Progressive Conditional Diffusion Models.

MVL-Loc: Leveraging Vision-Language Model for Generalizable Multi-Scene Camera Relocalization Advancing Pose-Guided Image Synthesis with Progressive Conditional Diffusion Models

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-06T19:51:52.756841Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:51:52.756841Z digest=sha256:f254a5e761bdc33d913eb46aa094410b718e10e59e45a07e6f0076b181782b52

Observation 64f9f7c1-362b-477c-a6a6-2db52a0c705f · outbound

This paper cites Boosting consistency in story visualization with rich-contextual conditional diffusion models,.

MVL-Loc: Leveraging Vision-Language Model for Generalizable Multi-Scene Camera Relocalization Boosting consistency in story visualization with rich-contextual conditional diffusion models,

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:51:53.933474Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-06T19:51:52.847978Z digest=sha256:1f4b8845eb571d32ea2a0834821c1c32f12cb68efcb9b7b13433e42013610fa0

Observation 2e3092c5-4a38-4a7d-a651-017588757960 · outbound

This paper cites 7-scenes dataset,.

MVL-Loc: Leveraging Vision-Language Model for Generalizable Multi-Scene Camera Relocalization 7-scenes dataset,

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:51:53.706874Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-06T19:51:52.928410Z digest=sha256:9689a3f07f064ea909ec3203c70974461dce3348112ac81a4f9bab4ce69cae7f

Observation cbb96d11-fce4-4cb5-9e62-0dc5cccb8b9c · outbound

This paper cites Do we really need scene-specific pose encoders?.

MVL-Loc: Leveraging Vision-Language Model for Generalizable Multi-Scene Camera Relocalization Do we really need scene-specific pose encoders?

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:51:53.521611Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-06T19:51:53.005097Z digest=sha256:490e92775a67daabac8ef07c8ef681fa781315d740097fb29a8d5b7f1f47653c

Observation d4c93214-bca1-4038-99db-956f9694ab88 · outbound

This paper cites Image-based localization using hourglass networks,.

MVL-Loc: Leveraging Vision-Language Model for Generalizable Multi-Scene Camera Relocalization Image-based localization using hourglass networks,

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:51:53.328502Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-06T19:51:53.074321Z digest=sha256:34d42c266873b37c5c086ceab4b7f093bf403b9d12df031ccaca6cee46c284fd

Pith citing papers

No inbound Pith citation observations are available.