Pith. sign in

Paper Citation Record · LEDGER

DialogueVPR: Towards Conversational Visual Place Recognition

As of 9 August 2026, this Paper Citation Record lists 55 of 55 outbound references and 0 inbound Pith citation observations for arXiv:2607.14115.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2607.14115 v1

Coverage vector

measured 55 of 55 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-02T14:39:43.408345Z

measured 55 of 55 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

55 of 55 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved55
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 6d81dd4d-3b32-48de-94ee-ab9cbbc455ec · outbound

This paper cites Gsv-cities: Toward appropriate supervised visual place recognition.Neurocomputing, 513:194–203, 2022.

DialogueVPR: Towards Conversational Visual Place Recognition Gsv-cities: Toward appropriate supervised visual place recognition.Neurocomputing, 513:194–203, 2022

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-02T14:39:39.143516Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T14:39:39.143516Z digest=sha256:1ae5aed1c76f1af755fbe01ec83f8a9195eabd99f768e1462326174962815c47

Observation df89369e-5249-4ecb-9a15-3459e7b03b51 · outbound

This paper cites Global Proxy-based Hard Mining for Visual Place Recognition.

DialogueVPR: Towards Conversational Visual Place Recognition Global Proxy-based Hard Mining for Visual Place Recognition

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-02T14:39:39.201606Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T14:39:39.201606Z digest=sha256:6cc0fb25ec366b94197b2180b880d04af104d9bdde5ac5af6ba64ef7850e597f

Observation f75dcdde-0338-43ac-b742-1a65ba9800d3 · outbound

This paper cites BoQ: A place is worth a bag of learnable queries.

DialogueVPR: Towards Conversational Visual Place Recognition BoQ: A place is worth a bag of learnable queries

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-02T14:39:39.267722Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T14:39:39.267722Z digest=sha256:dd90f93d63d60a67184d575671e044e6667b35e3f32a88c2e6040fd54f8b9ce1

Observation 044524bf-f53e-4a64-ade6-f1f1d0c20243 · outbound

This paper cites Netvlad: Cnn architecture for weakly supervised place recognition.

DialogueVPR: Towards Conversational Visual Place Recognition Netvlad: Cnn architecture for weakly supervised place recognition

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-02T14:39:39.347228Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T14:39:39.347228Z digest=sha256:98827db4de3374942e17a4f54fa68a086cb899cab55b63b8c33f663b9c864d25

Observation 9e50e462-1665-41f8-b7c2-53445175215a · outbound

This paper cites Ask&confirm: active detail enriching for cross-modal retrieval with partial query.

DialogueVPR: Towards Conversational Visual Place Recognition Ask&confirm: active detail enriching for cross-modal retrieval with partial query

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-02T14:39:39.430324Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T14:39:39.430324Z digest=sha256:f1def5ef158abbcbf81d1085670c876dc659324a34e35bd2d51dd5977127c838

Observation 72b0da76-59b1-42e9-9015-c994ff59c504 · outbound

This paper cites where am i?.

DialogueVPR: Towards Conversational Visual Place Recognition where am i?

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-02T14:39:39.508679Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T14:39:39.508679Z digest=sha256:1a7907e134524f19b6d3b76268e4a383b61f28d8f91c1c9ce4d5d3a01646038a

Observation 94fb2119-50ca-418a-8717-1a8bb761545d · outbound

This paper cites Region Matters: Efficient and Reliable Region-Aware Visual Place Recognition.

DialogueVPR: Towards Conversational Visual Place Recognition Region Matters: Efficient and Reliable Region-Aware Visual Place Recognition

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-02T14:39:39.583709Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T14:39:39.583709Z digest=sha256:a4938c9fd6b68367af3e0e25025d453ad05f48c42a3ac3c17c06af191020f501

Observation 5d6f3a76-388d-452b-a6c2-683dc08e31e0 · outbound

This paper cites SAGE: Spatial-visual adaptive graph exploration for efficient visual place recognition.

DialogueVPR: Towards Conversational Visual Place Recognition SAGE: Spatial-visual adaptive graph exploration for efficient visual place recognition

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-02T14:39:39.665070Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T14:39:39.665070Z digest=sha256:f767d001d3cca36272968ef3d881376671b2f55750152f4ebbbd0857c1f92321

Observation 545e630d-aa1a-410a-89e3-1b0b5d0f2c74 · outbound

This paper cites Towards natural language-guided drones: Geotext-1652 benchmark with spatial relation matching.

DialogueVPR: Towards Conversational Visual Place Recognition Towards natural language-guided drones: Geotext-1652 benchmark with spatial relation matching

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-02T14:39:39.747412Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T14:39:39.747412Z digest=sha256:4f84121af4a60bae82219fc3c43e6b81764305feae302a86774c04aee0838b77

Observation f47764ab-4990-4e13-a092-3019e1615f1c · outbound

This paper cites Sam- ple4geo: Hard negative sampling for cross-view geo- localisation.

DialogueVPR: Towards Conversational Visual Place Recognition Sam- ple4geo: Hard negative sampling for cross-view geo- localisation

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-02T14:39:39.797523Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T14:39:39.797523Z digest=sha256:13043f200605326d613b4d53931e494061baa4a8d686b6efb05f50e775d6acde

Observation 0a48b0d7-f127-4d49-a3e8-f94698d8d2c1 · outbound

This paper cites Seq- matchnet: Contrastive learning with sequence matching for place recognition & relocalization.

DialogueVPR: Towards Conversational Visual Place Recognition Seq- matchnet: Contrastive learning with sequence matching for place recognition & relocalization

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-02T14:39:39.874044Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T14:39:39.874044Z digest=sha256:f1a605e4817ff7a347be4953a92d5d6ff27b68ed8984067ca133ab8fe27a5ec5

Observation eef4fbc7-5dbf-47d1-91c0-f4b1da6bebb9 · outbound

This paper cites DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning.

DialogueVPR: Towards Conversational Visual Place Recognition DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-02T14:39:39.959139Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T14:39:39.959139Z digest=sha256:f7fc3f5fb77aeeedc1c3ead54d6ea8b697397002f6a08cc9f463c7a5834ed521

Observation a42aebdc-2615-4502-b72a-d82548e8153e · outbound

This paper cites Textplace: Visual place recognition and topolog- ical localization through reading scene texts.

DialogueVPR: Towards Conversational Visual Place Recognition Textplace: Visual place recognition and topolog- ical localization through reading scene texts

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-02T14:39:40.020449Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T14:39:40.020449Z digest=sha256:161f23f83272e55b38a4b16d7dada7aa77a3f2d8f1c890f92d6a4fe0f6583418

Observation 4690cd53-8964-4e74-9e52-c19fef2cc5ba · outbound

This paper cites Progeo: Generating prompts through image- text contrastive learning for visual geo-localization.

DialogueVPR: Towards Conversational Visual Place Recognition Progeo: Generating prompts through image- text contrastive learning for visual geo-localization

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-02T14:39:40.116057Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T14:39:40.116057Z digest=sha256:ae720d828a7a324c6cc353ee6604bf97af791fedfda666224a2fe43b2c43db31

Observation 446b4d00-f85e-4c5d-8eb9-c87a75675a95 · outbound

This paper cites Optimal transport ag- gregation for visual place recognition.

DialogueVPR: Towards Conversational Visual Place Recognition Optimal transport ag- gregation for visual place recognition

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-02T14:39:40.196754Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T14:39:40.196754Z digest=sha256:c37e0fb4f8f053efdb3b781888e0c1f8ab199a62730d8995a704ce384e704d1b

Observation 76ae7ba9-544e-4d14-ab45-24c25940845e · outbound

This paper cites Close, but not there: Boosting geographic distance sensitivity in visual place recognition.

DialogueVPR: Towards Conversational Visual Place Recognition Close, but not there: Boosting geographic distance sensitivity in visual place recognition

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-02T14:39:40.282316Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T14:39:40.282316Z digest=sha256:4f1a04d01e6c1442defd99b447e4086feee7d518119724817e50c5f6ed49d954

Observation 6c0923c1-d8e3-4499-8287-094c29b5ca15 · outbound

This paper cites OpenAI o1 System Card.

DialogueVPR: Towards Conversational Visual Place Recognition OpenAI o1 System Card

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-02T14:39:40.364752Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T14:39:40.364752Z digest=sha256:d8c758301754278514f40154d1063dafdf2c60005d3bee28f47978c63395df3d

Observation 7683195f-e8a4-4dd5-8b44-61b8c4732fe3 · outbound

This paper cites Cross-modal implicit relation rea- soning and aligning for text-to-image person retrieval.

DialogueVPR: Towards Conversational Visual Place Recognition Cross-modal implicit relation rea- soning and aligning for text-to-image person retrieval

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-02T14:39:40.447356Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T14:39:40.447356Z digest=sha256:eea66da54c9b61218b2213808f32482f60dc3b7533607dce04ae807c156bfbf6

Observation 9aaf1dc2-bcc0-488b-a4cb-850d328f9982 · outbound

This paper cites Text2pos: Text-to-point-cloud cross-modal localiza- tion.

DialogueVPR: Towards Conversational Visual Place Recognition Text2pos: Text-to-point-cloud cross-modal localiza- tion

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-02T14:39:40.507494Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T14:39:40.507494Z digest=sha256:8c72011910ecd0a9b63285a38ef622813e995064200279e4b31ddafcef32bc35

Observation 2c67acc5-ddc2-4a7e-9d4e-241198cfcbf3 · outbound

This paper cites Whittlesearch: Interactive image search with relative at- tribute feedback.International Journal of Computer Vision, 115(2):185–210, 2015.

DialogueVPR: Towards Conversational Visual Place Recognition Whittlesearch: Interactive image search with relative at- tribute feedback.International Journal of Computer Vision, 115(2):185–210, 2015

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-02T14:39:40.587880Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T14:39:40.587880Z digest=sha256:219a83588120fb3424ea3e96f548c47e4dc579b0627c634eb1ccde76afafbdb6

Observation 2df1e1eb-36b9-4c46-ad15-308003c429bc · outbound

This paper cites Cosmo: Content-style modulation for image retrieval with text feed- back.

DialogueVPR: Towards Conversational Visual Place Recognition Cosmo: Content-style modulation for image retrieval with text feed- back

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-02T14:39:40.698349Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T14:39:40.698349Z digest=sha256:7edb64bd7ba3ba342d5c675ebe10de26b9231540b7cfd8ce055dafdb32f2db05

Observation 5958a955-1502-423d-9104-534c0691dc9a · outbound

This paper cites Interactive Text-to-Image Retrieval with Large Language Models: A Plug-and-Play Approach.

DialogueVPR: Towards Conversational Visual Place Recognition Interactive Text-to-Image Retrieval with Large Language Models: A Plug-and-Play Approach

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-02T14:39:40.780873Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T14:39:40.780873Z digest=sha256:505c112119b25cdebf706dd9e261de74052afe884ac886baaafb5f3ca2b9d3cc

Observation adc00c94-d481-4aae-8ec3-ba0451eb6b98 · outbound

This paper cites Chatting makes perfect: Chat-based image retrieval.

DialogueVPR: Towards Conversational Visual Place Recognition Chatting makes perfect: Chat-based image retrieval

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-02T14:39:40.841666Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T14:39:40.841666Z digest=sha256:5cb53d2b61d1db646b24534f45eeb269720c512505cbcbe50095eb4d67b86bbd

Observation 680708bc-0ab2-47cb-bec3-3d564a0c4b7a · outbound

This paper cites Omnicity: Omnipotent city understanding with multi-level and multi- view images.

DialogueVPR: Towards Conversational Visual Place Recognition Omnicity: Omnipotent city understanding with multi-level and multi- view images

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-02T14:39:40.918209Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T14:39:40.918209Z digest=sha256:58e9e2228153e868596576ec11950ab83e4faddc23b6e630df7d89b965c18792

Observation 45506850-506e-4812-a644-e754d0b80a59 · outbound

This paper cites Simple baselines for interactive video retrieval with questions and answers.

DialogueVPR: Towards Conversational Visual Place Recognition Simple baselines for interactive video retrieval with questions and answers

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-02T14:39:41.001765Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T14:39:41.001765Z digest=sha256:4ae53a4b424399b95649ac2f1ceb9d1149975c477ed90753015fd7bc1ff1ea42

Observation 1ae29204-3fda-4586-a3b8-5659dc0e2a5c · outbound

This paper cites Visual-RFT: Visual Reinforcement Fine-Tuning.

DialogueVPR: Towards Conversational Visual Place Recognition Visual-RFT: Visual Reinforcement Fine-Tuning

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-02T14:39:41.107742Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T14:39:41.107742Z digest=sha256:13ca319a485a88f04fbc027a3d495181ee173443c3a940eb0aeb5b3adab7fc80

Observation 7b73d86a-4423-482b-ba41-2294a4241b23 · outbound

This paper cites Towards seamless adapta- tion of pre-trained models for visual place recognition.

DialogueVPR: Towards Conversational Visual Place Recognition Towards seamless adapta- tion of pre-trained models for visual place recognition

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-02T14:39:41.177835Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T14:39:41.177835Z digest=sha256:b5437841ee6d6658a7a57199fbd1abe2cb24c614e3fc34a5c04187dbcd9ae6c7

Observation 3bc298de-c8fe-4ffb-9dbe-8db09f7956cc · outbound

This paper cites Supervlad: Com- pact and robust image descriptors for visual place recogni- tion.Advances in Neural Information Processing Systems, 37:5789–5816, 2024.

DialogueVPR: Towards Conversational Visual Place Recognition Supervlad: Com- pact and robust image descriptors for visual place recogni- tion.Advances in Neural Information Processing Systems, 37:5789–5816, 2024

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-02T14:39:41.246535Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T14:39:41.246535Z digest=sha256:943235e238dca5b90e126a804a20ba1e8c17b7b61d8a84a9c7a394e7738fc999

Observation c8f66496-45c6-4bed-b1c4-ae672778d6e8 · outbound

This paper cites LLaVA-ReID: Selective Multi-image Questioner for Interactive Person Re-Identification.

DialogueVPR: Towards Conversational Visual Place Recognition LLaVA-ReID: Selective Multi-image Questioner for Interactive Person Re-Identification

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-02T14:39:41.329868Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T14:39:41.329868Z digest=sha256:1b3b7d34d95b0a4b110b5562c2fe75dfb3d73f09e49638e8ab474359bc22e508

Observation ab159d0e-7c40-4d45-bc96-a713fd345018 · outbound

This paper cites Tell Me Where You Are: Multimodal LLMs Meet Place Recognition.

DialogueVPR: Towards Conversational Visual Place Recognition Tell Me Where You Are: Multimodal LLMs Meet Place Recognition

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-02T14:39:41.410192Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T14:39:41.410192Z digest=sha256:c2463fb37baaa3db1b4e23b6fa471f28c6a97caea1b71ba3dc079c5b80fe666a

Observation b9f26f10-e957-456c-9a85-2db65a7fc899 · outbound

This paper cites Learn- ing to retrieve videos by asking questions.

DialogueVPR: Towards Conversational Visual Place Recognition Learn- ing to retrieve videos by asking questions

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-02T14:39:41.491749Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T14:39:41.491749Z digest=sha256:1d5aa73ea9007c79e85d8c06fcdc863ef4390b0a7f4d150e1e6a3ca35cb2c35c

Observation 883543fb-d6fb-4b9e-90f3-75cba2a29af0 · outbound

This paper cites Emvp: Em- bracing visual foundation model for visual place recognition with centroid-free probing.Advances in Neural Information Processing Systems, 37:120928–120950, 2024.

DialogueVPR: Towards Conversational Visual Place Recognition Emvp: Em- bracing visual foundation model for visual place recognition with centroid-free probing.Advances in Neural Information Processing Systems, 37:120928–120950, 2024

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-02T14:39:41.577075Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T14:39:41.577075Z digest=sha256:cf4d56fcd51426a27f775872f5ec8b22ebc9ac849bef2b65b369accface4f335

Observation 70250d35-1285-4c4d-adab-0bcfc9f16dd5 · outbound

This paper cites Learning transferable visual models from natural language supervi- sion.

DialogueVPR: Towards Conversational Visual Place Recognition Learning transferable visual models from natural language supervi- sion

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-02T14:39:41.660639Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T14:39:41.660639Z digest=sha256:9cc52c96d0775556eb736640ef64b35f9e6abb0ac84cdff2fb561b39cb58efe7

Observation 57a845fa-a392-480e-8ca0-f5458d828766 · outbound

This paper cites From coarse to fine: Robust hierarchical localization at large scale.

DialogueVPR: Towards Conversational Visual Place Recognition From coarse to fine: Robust hierarchical localization at large scale

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-02T14:39:41.746035Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T14:39:41.746035Z digest=sha256:75f4978e5654294e3c1f502000bf86fa0bb25742250e0ed317ccd824c87bf4fc

Observation 4cd26d6c-3e2e-4fe5-a3cf-8eaf757650ce · outbound

This paper cites Bridging Text and Vision: A Multi-View Text-Vision Registration Approach for Cross-Modal Place Recognition.

DialogueVPR: Towards Conversational Visual Place Recognition Bridging Text and Vision: A Multi-View Text-Vision Registration Approach for Cross-Modal Place Recognition

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-02T14:39:41.830609Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T14:39:41.830609Z digest=sha256:200e9b8ff6ff9232f382ff7e87f65ef858aa88a78be88e8e49d3965b6f4f6404

Observation 9a5d4326-f24f-40d1-912b-62e2abfab7b6 · outbound

This paper cites DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models.

DialogueVPR: Towards Conversational Visual Place Recognition DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-02T14:39:41.913539Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T14:39:41.913539Z digest=sha256:071536005f328732d3d6f9acafa62d87ef41f09e29022a721c8b6fa191d2ce73

Observation 11454ed3-1d24-4970-9d84-6935fb79c2c9 · outbound

This paper cites VLM-R1: A Stable and Generalizable R1-style Large Vision-Language Model.

DialogueVPR: Towards Conversational Visual Place Recognition VLM-R1: A Stable and Generalizable R1-style Large Vision-Language Model

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-02T14:39:41.971750Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T14:39:41.971750Z digest=sha256:1a8260e3b39cb2c2e8a36cef86619ae1a4b72f2c1a3011bc8c62c90e84d63bf2

Observation 49a055f8-a119-403b-9ee0-899d7e21b6fd · outbound

This paper cites Where am i looking at? joint location and orientation es- timation by cross-view matching.

DialogueVPR: Towards Conversational Visual Place Recognition Where am i looking at? joint location and orientation es- timation by cross-view matching

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-02T14:39:42.048832Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T14:39:42.048832Z digest=sha256:65e545a2204f8b7aee62748fad786d2721ef8a667d40c4e938a9d41c934a7247

Observation 0095fdb1-8ec0-4bfe-bcd1-05de325cb63e · outbound

This paper cites TextInPlace: Indoor Visual Place Recognition in Repetitive Structures with Scene Text Spotting and Verification.

DialogueVPR: Towards Conversational Visual Place Recognition TextInPlace: Indoor Visual Place Recognition in Repetitive Structures with Scene Text Spotting and Verification

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-02T14:39:42.108448Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T14:39:42.108448Z digest=sha256:e77360c36c4ae3a14a83f4561c459abea89ede46d162543aefa5106c7d39f801

Observation d8be864d-3b63-48a8-8919-b2aee52037bb · outbound

This paper cites Loc4plan: Locating be- fore planning for outdoor vision and language navigation.

DialogueVPR: Towards Conversational Visual Place Recognition Loc4plan: Locating be- fore planning for outdoor vision and language navigation

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-02T14:39:42.159054Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T14:39:42.159054Z digest=sha256:2554470ba9269309e69d0389f7db82b4b73b8140c831e8e6c1c1906fa74feabd

Observation cd8e3e24-d008-4bdc-a0eb-7760d5fa08a0 · outbound

This paper cites Focus on local: Finding reliable discriminative regions for visual place recognition.

DialogueVPR: Towards Conversational Visual Place Recognition Focus on local: Finding reliable discriminative regions for visual place recognition

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-02T14:39:42.243908Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T14:39:42.243908Z digest=sha256:50ef7978f70fbf1dc315e81891ee60d88e2caf43e8efa5c92b6d6438cdb6bc86

Observation 9b832fc1-88de-495f-a0cb-f9419250150f · outbound

This paper cites LVLM-empowered Multi-modal Representation Learning for Visual Place Recognition.

DialogueVPR: Towards Conversational Visual Place Recognition LVLM-empowered Multi-modal Representation Learning for Visual Place Recognition

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-02T14:39:42.328089Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T14:39:42.328089Z digest=sha256:8898c02ff1177247182f130ee26867edc8c774b218f4ed21bb3aad3fccfc2756

Observation 3aecf139-7d29-482f-92d1-f270b7386029 · outbound

This paper cites Multi-similarity loss with general pair weighting for deep metric learning.

DialogueVPR: Towards Conversational Visual Place Recognition Multi-similarity loss with general pair weighting for deep metric learning

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-02T14:39:42.410690Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T14:39:42.410690Z digest=sha256:33e5386bb6b130f625517efc2b8cb3cd3a52806a6c690c94bce0f198de3344ec

Observation 01364525-4486-431f-a994-af26f6e2aae0 · outbound

This paper cites Fine-grained cross-view geo-localization using a correlation-aware homography estimator.Advances in Neu- ral Information Processing Systems, 36:5301–5319, 2023.

DialogueVPR: Towards Conversational Visual Place Recognition Fine-grained cross-view geo-localization using a correlation-aware homography estimator.Advances in Neu- ral Information Processing Systems, 36:5301–5319, 2023

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-02T14:39:42.497204Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T14:39:42.497204Z digest=sha256:ae4fbdf84c3ba90e9477b19e6096efe0820216f956743532dca40bd7e611d518

Observation 2fd26477-f525-4538-8146-8da39567677a · outbound

This paper cites A theoretical analysis of ndcg type ranking measures.

DialogueVPR: Towards Conversational Visual Place Recognition A theoretical analysis of ndcg type ranking measures

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-02T14:39:42.610369Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T14:39:42.610369Z digest=sha256:f7393069ade6c9c01e616253c73488b78a6aed26c9ce3b7299d1ad417adb02ac

Observation aa3d28c3-fd8f-4220-9fa5-354aa6b04fac · outbound

This paper cites Text2loc: 3d point cloud localization from natural language.

DialogueVPR: Towards Conversational Visual Place Recognition Text2loc: 3d point cloud localization from natural language

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-02T14:39:42.672422Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T14:39:42.672422Z digest=sha256:cd904a9b39c2be82320aaba5690b4f16d64b37ff054717e4a9c58eac58ff5ed3

Observation 88f6c4bb-f38d-4169-9e13-dd79b18fabbe · outbound

This paper cites Adapting fine-grained cross-view localization to areas with- out fine ground truth.

DialogueVPR: Towards Conversational Visual Place Recognition Adapting fine-grained cross-view localization to areas with- out fine ground truth

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-02T14:39:42.753892Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T14:39:42.753892Z digest=sha256:1d1395bcf9cc1bb7447b15ef7e50201f777a6ebb7ed5b4114632f913598bbd5e

Observation cead20c2-a77e-4e44-bd1b-dbd3faf29979 · outbound

This paper cites Flair: Vlm with fine- grained language-informed image representations.

DialogueVPR: Towards Conversational Visual Place Recognition Flair: Vlm with fine- grained language-informed image representations

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-02T14:39:42.842868Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T14:39:42.842868Z digest=sha256:567c977f2a3532175a3a2e6ba4e20f372cacc5ef9c4b85bca4fc562c10b648c6

Observation 8addccce-036f-48d6-98bb-def375b1dbfc · outbound

This paper cites FG-CLIP: Fine-Grained Visual and Textual Alignment.

DialogueVPR: Towards Conversational Visual Place Recognition FG-CLIP: Fine-Grained Visual and Textual Alignment

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-02T14:39:42.933706Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T14:39:42.933706Z digest=sha256:e13a88bc0ca1a39ef9affd707e319abe948f24651a72f06625f245bd030a95da

Observation 0bf30321-753e-429e-b6ed-9a7eb5f6c2f9 · outbound

This paper cites Skydiffusion: Street-to-satellite im- age synthesis with diffusion models and bev paradigm.arXiv e-prints, pages arXiv–2408, 2024.

DialogueVPR: Towards Conversational Visual Place Recognition Skydiffusion: Street-to-satellite im- age synthesis with diffusion models and bev paradigm.arXiv e-prints, pages arXiv–2408, 2024

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-02T14:39:43.019303Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T14:39:43.019303Z digest=sha256:626a8fa9511e78ce0914946f33f955b76f22c48eac79d815e7da79ef4073697d

Observation 793e0ed6-328c-4e59-bb67-885edfa4bce5 · outbound

This paper cites Cross-view image geo- localization with panorama-bev co-retrieval network.

DialogueVPR: Towards Conversational Visual Place Recognition Cross-view image geo- localization with panorama-bev co-retrieval network

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-02T14:39:43.109120Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T14:39:43.109120Z digest=sha256:ed8ec262c3de9bc705305c3e7930cfaebce9c131bc633663542adcd937acc1ff

Observation c01cdd01-a29b-4220-919f-b426f732598e · outbound

This paper cites Where am i? cross-view geo-localization with natural language descrip- tions.

DialogueVPR: Towards Conversational Visual Place Recognition Where am i? cross-view geo-localization with natural language descrip- tions

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-02T14:39:43.183346Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T14:39:43.183346Z digest=sha256:9d3397d76b83277558b57fef4d1fc01d48780913809d90115fb818257fd59a9a

Observation 21a01e68-f8b5-4ade-942a-3f48c3c321c7 · outbound

This paper cites Long-clip: Unlocking the long-text capability of clip.

DialogueVPR: Towards Conversational Visual Place Recognition Long-clip: Unlocking the long-text capability of clip

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-02T14:39:43.281595Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T14:39:43.281595Z digest=sha256:3248b2ac1988337662e5203fde63bf9ad0cc6108ad23a54911f104e18c29d3d2

Observation 98a8c816-ee0c-4969-9424-a2cb2a7d18b7 · outbound

This paper cites NaVid: Video-based VLM Plans the Next Step for Vision-and-Language Navigation.

DialogueVPR: Towards Conversational Visual Place Recognition NaVid: Video-based VLM Plans the Next Step for Vision-and-Language Navigation

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-02T14:39:43.349828Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T14:39:43.349828Z digest=sha256:a06d1bd68fec7214f2134c8958eedf4a491067170ca25eecb3d5682f2ca57ef2

Observation 9695888d-038a-4ca4-bb40-6d3f815fdec8 · outbound

This paper cites Enhancing interactive image retrieval with query rewriting using large language models and vision lan- guage models.

DialogueVPR: Towards Conversational Visual Place Recognition Enhancing interactive image retrieval with query rewriting using large language models and vision lan- guage models

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-02T14:39:43.408345Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T14:39:43.408345Z digest=sha256:31a795d2b3bd5902e513853eea004956d5ea4d97fc26e354160d87cc192125e7

Pith citing papers

No inbound Pith citation observations are available.