Pith. sign in

Paper Citation Record · LEDGER

Scale, Don't Fine-tune: Guiding Multimodal LLMs for Efficient Visual Place Recognition at Test-Time

As of 10 August 2026, this Paper Citation Record lists 28 of 28 outbound references and 0 inbound Pith citation observations for arXiv:2509.02129.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2509.02129 v1

Coverage vector

measured 28 of 28 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-05T11:57:17.255005Z

measured 28 of 28 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-10T06:31:04.303077+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

28 of 28 outbound references displayed

  • verified exact2
  • verified fuzzy0
  • unresolved26
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation bf1d124e-84b7-41ae-8bdc-4010ad07f9d6 · outbound

This paper cites an unresolved cited work.

Scale, Don't Fine-tune: Guiding Multimodal LLMs for Efficient Visual Place Recognition at Test-Time Unresolved cited work

Reference 1

Resolution
unresolved
raw_fallback, observed 2026-08-05T11:57:17.877913Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-05T11:57:16.316998Z digest=sha256:fbc120cde37ed1c956d2b60850310ab75119f81c2ac7b73c448dd9372b7183a5

Observation f5110ec7-93fd-4714-b01c-f7e3064378ef · outbound

This paper cites an unresolved cited work.

Scale, Don't Fine-tune: Guiding Multimodal LLMs for Efficient Visual Place Recognition at Test-Time Unresolved cited work

Reference 2

Resolution
unresolved
raw_fallback, observed 2026-08-05T11:57:17.864379Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-05T11:57:16.355902Z digest=sha256:6fbc90962116cdf5c5aae4e1df6058777e59b81045b559a378edecaf48291b50

Observation fd67a794-0360-486f-a8e0-5680d2dda27f · outbound

This paper cites an unresolved cited work.

Scale, Don't Fine-tune: Guiding Multimodal LLMs for Efficient Visual Place Recognition at Test-Time Unresolved cited work

Reference 3

Resolution
unresolved
raw_fallback, observed 2026-08-05T11:57:17.850435Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-05T11:57:16.455330Z digest=sha256:0a259d195edce18d06d9c356a185935168819402e4f0688234ffac8a0b056d55

Observation 5990be6f-6b82-403f-a1be-9c77f2b4f35c · outbound

This paper cites an unresolved cited work.

Scale, Don't Fine-tune: Guiding Multimodal LLMs for Efficient Visual Place Recognition at Test-Time Unresolved cited work

Reference 4

Resolution
unresolved
raw_fallback, observed 2026-08-05T11:57:17.836818Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-05T11:57:16.519091Z digest=sha256:8417897a53c41b22e79b5144a86e286ed855a31b03a3e2e0b63bd0b55a908f5e

Observation e029c155-342d-4b6c-a22b-0cc2dc634b09 · outbound

This paper cites an unresolved cited work.

Scale, Don't Fine-tune: Guiding Multimodal LLMs for Efficient Visual Place Recognition at Test-Time Unresolved cited work

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-05T11:57:16.596950Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T11:57:16.596950Z digest=sha256:1c4f699989b6bd8ed76ecdd4132c7f0c33668399e16e5c33671dd470788cb6b0

Observation 87999208-3bcf-46a6-90d6-3171d90122ce · outbound

This paper cites an unresolved cited work.

Scale, Don't Fine-tune: Guiding Multimodal LLMs for Efficient Visual Place Recognition at Test-Time Unresolved cited work

Reference 6

Resolution
unresolved
raw_fallback, observed 2026-08-05T11:57:17.822687Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-05T11:57:16.683160Z digest=sha256:bfe69cc6623c16ca9d609123bc0f8117d874b803d55b4963fee45be13b5519d1

Observation 14285457-87e1-4d05-a671-243df4b6e228 · outbound

This paper cites Alphazero-like Tree-Search can Guide Large Language Model Decoding and Training.

Scale, Don't Fine-tune: Guiding Multimodal LLMs for Efficient Visual Place Recognition at Test-Time Alphazero-like Tree-Search can Guide Large Language Model Decoding and Training

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-05T11:57:16.716582Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T11:57:16.716582Z digest=sha256:522037f8398de52efcd93beeaf1fb73eff8a7278681d6bb0cc1d21b9c977da37

Observation 0cff251c-719f-4a14-9b30-ce93e8ae5b13 · outbound

This paper cites an unresolved cited work.

Scale, Don't Fine-tune: Guiding Multimodal LLMs for Efficient Visual Place Recognition at Test-Time Unresolved cited work

Reference 8

Resolution
unresolved
raw_fallback, observed 2026-08-05T11:57:17.808279Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-05T11:57:16.800663Z digest=sha256:1dcc46840320ced48931f2b954fc1dcd0ae52260c1848deb490fca97b85eca19

Observation bcf3f1b3-a73c-42a9-be4c-295b3440c589 · outbound

This paper cites an unresolved cited work.

Scale, Don't Fine-tune: Guiding Multimodal LLMs for Efficient Visual Place Recognition at Test-Time Unresolved cited work

Reference 9

Resolution
unresolved
raw_fallback, observed 2026-08-05T11:57:17.794198Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-05T11:57:16.955019Z digest=sha256:22d9d0c0684cad7998f49cfd445f445d12453805ddc3a1d89d8179a4884eafd8

Observation 2b2e8ff4-0239-490e-9170-fb08871dc77f · outbound

This paper cites an unresolved cited work.

Scale, Don't Fine-tune: Guiding Multimodal LLMs for Efficient Visual Place Recognition at Test-Time Unresolved cited work

Reference 10

Resolution
unresolved
raw_fallback, observed 2026-08-05T11:57:17.779771Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-05T11:57:17.116224Z digest=sha256:2fbf95f3ceb3fbb55a593171ab508aa94cf4012a00d35b9573de27617532a354

Observation a17a134d-2444-482f-89be-826632a77ee2 · outbound

This paper cites an unresolved cited work.

Scale, Don't Fine-tune: Guiding Multimodal LLMs for Efficient Visual Place Recognition at Test-Time Unresolved cited work

Reference 11

Resolution
unresolved
raw_fallback, observed 2026-08-05T11:57:17.764029Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-05T11:57:17.120543Z digest=sha256:79dbd288c587333bb22a13d78feafa7441b7ca6df722375e9ecbd14777878077

Observation 6b92c121-f72f-463a-919a-55f1ef69da0f · outbound

This paper cites an unresolved cited work.

Scale, Don't Fine-tune: Guiding Multimodal LLMs for Efficient Visual Place Recognition at Test-Time Unresolved cited work

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-05T11:57:17.185085Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T11:57:17.185085Z digest=sha256:9a51b3a71a517288bdd8489b9e161472dbbf41226d1b773fec5821630acdd359

Observation 30b25eca-efca-4649-a6c8-d395812f596d · outbound

This paper cites Tell Me Where You Are: Multimodal LLMs Meet Place Recognition.

Scale, Don't Fine-tune: Guiding Multimodal LLMs for Efficient Visual Place Recognition at Test-Time Tell Me Where You Are: Multimodal LLMs Meet Place Recognition

Reference 13

Resolution
verified exact
local_arxiv, observed 2026-08-05T11:57:17.435287Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-05T11:57:17.189338Z digest=sha256:f8fe06b68a68f6c8d889523ca09a3e2ece8b459be2cfc89cfef6f5e5d32ebffb

Observation e23b4d26-4498-4661-bff8-d1c5facd7bdb · outbound

This paper cites an unresolved cited work.

Scale, Don't Fine-tune: Guiding Multimodal LLMs for Efficient Visual Place Recognition at Test-Time Unresolved cited work

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-05T11:57:17.194612Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T11:57:17.194612Z digest=sha256:fe55061fb820849636634f5f6f43820c4c8854643e10e50cafdb7d775ab1f23b

Observation c266c17b-4bbe-4c14-a8b9-b450801b6029 · outbound

This paper cites TreeBoN: Enhancing Inference-Time Alignment with Speculative Tree-Search and Best-of-N Sampling.

Scale, Don't Fine-tune: Guiding Multimodal LLMs for Efficient Visual Place Recognition at Test-Time TreeBoN: Enhancing Inference-Time Alignment with Speculative Tree-Search and Best-of-N Sampling

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-05T11:57:17.198660Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T11:57:17.198660Z digest=sha256:016b471e70ddc9c33351b17fee8e744432b81b34ae4ddc52565f9b5832124f1e

Observation 509ee550-866c-4896-8ebe-960f1fe8c9d4 · outbound

This paper cites Scaling LLM Test-Time Compute Optimally can be More Effective than Scaling Model Parameters.

Scale, Don't Fine-tune: Guiding Multimodal LLMs for Efficient Visual Place Recognition at Test-Time Scaling LLM Test-Time Compute Optimally can be More Effective than Scaling Model Parameters

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-05T11:57:17.203309Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T11:57:17.203309Z digest=sha256:a704c88a875cf302cab99359d1310710db655faee136fadc9978e08a570963fb

Observation 04a8acd0-7b26-4efa-854d-20f1bf1a13b3 · outbound

This paper cites an unresolved cited work.

Scale, Don't Fine-tune: Guiding Multimodal LLMs for Efficient Visual Place Recognition at Test-Time Unresolved cited work

Reference 17

Resolution
unresolved
raw_fallback, observed 2026-08-05T11:57:17.749532Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-05T11:57:17.208210Z digest=sha256:e91ec9c9da33659863a4dbd58239b8577f55b6837c366d4f61656914d02066a2

Observation f9a92daa-1f23-4d7f-ac87-7567485d0efa · outbound

This paper cites an unresolved cited work.

Scale, Don't Fine-tune: Guiding Multimodal LLMs for Efficient Visual Place Recognition at Test-Time Unresolved cited work

Reference 18

Resolution
unresolved
raw_fallback, observed 2026-08-05T11:57:17.734489Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-05T11:57:17.212218Z digest=sha256:30731d0971c5b926fbd84f90294fbd0ce80d12e2e46aa384aada158ba1825548

Observation 65273ceb-6a4d-4eb7-ab27-4ff2ade810b1 · outbound

This paper cites an unresolved cited work.

Scale, Don't Fine-tune: Guiding Multimodal LLMs for Efficient Visual Place Recognition at Test-Time Unresolved cited work

Reference 19

Resolution
unresolved
raw_fallback, observed 2026-08-05T11:57:17.720097Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-05T11:57:17.216064Z digest=sha256:7655a53935ae873ab7995d8c71ff5887b203ee67a860a7ad4c13b5ad0f27c775

Observation 324a6c31-86b4-420a-888f-0b02322b2b8e · outbound

This paper cites an unresolved cited work.

Scale, Don't Fine-tune: Guiding Multimodal LLMs for Efficient Visual Place Recognition at Test-Time Unresolved cited work

Reference 20

Resolution
unresolved
raw_fallback, observed 2026-08-05T11:57:17.705310Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-05T11:57:17.220188Z digest=sha256:369778a25f943713762d5a81d0373b07c4249e1e6e1ddddec47d1fedf8f66cbb

Observation 1d75041c-d04c-407a-a219-1a13e8945aee · outbound

This paper cites an unresolved cited work.

Scale, Don't Fine-tune: Guiding Multimodal LLMs for Efficient Visual Place Recognition at Test-Time Unresolved cited work

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-05T11:57:17.224316Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T11:57:17.224316Z digest=sha256:276e4374de3fd97fcd0012c9d29d1ddf5be49e2f8220a0c46cc2a1931d864ad4

Observation 60d893ea-1597-4763-9791-91d1b4f62f40 · outbound

This paper cites an unresolved cited work.

Scale, Don't Fine-tune: Guiding Multimodal LLMs for Efficient Visual Place Recognition at Test-Time Unresolved cited work

Reference 22

Resolution
unresolved
raw_fallback, observed 2026-08-05T11:57:17.680668Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-05T11:57:17.228965Z digest=sha256:81d5fa8985037d3382e0325d38d78d9def80333a2389530df4582a4c03eb4c1d

Observation 40a982e5-ef36-4df9-9b57-8794f38db75d · outbound

This paper cites an unresolved cited work.

Scale, Don't Fine-tune: Guiding Multimodal LLMs for Efficient Visual Place Recognition at Test-Time Unresolved cited work

Reference 23

Resolution
unresolved
raw_fallback, observed 2026-08-05T11:57:17.666062Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-05T11:57:17.232985Z digest=sha256:b1cbd33f2bcab5435f8c680b0933286b8b18b0eb4593a0fa532241ed4516f98d

Observation 420fffb0-6471-4d2c-9441-8ffda85c88f6 · outbound

This paper cites an unresolved cited work.

Scale, Don't Fine-tune: Guiding Multimodal LLMs for Efficient Visual Place Recognition at Test-Time Unresolved cited work

Reference 24

Resolution
unresolved
raw_fallback, observed 2026-08-05T11:57:17.651568Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-05T11:57:17.237103Z digest=sha256:3872915e0f39ac25dba5fa8895bb3e530ba0eddb588f0cd760a0ac68c9bed5dc

Observation 1e339af1-e62b-49c2-8f8b-6aa90c937897 · outbound

This paper cites an unresolved cited work.

Scale, Don't Fine-tune: Guiding Multimodal LLMs for Efficient Visual Place Recognition at Test-Time Unresolved cited work

Reference 25

Resolution
unresolved
raw_fallback, observed 2026-08-05T11:57:17.636106Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-05T11:57:17.241091Z digest=sha256:b41900d2e7d3b4d1f8d61b5fc4341af47450b6c5cef422caf35059b894506f38

Observation ee72edad-a407-4f03-a9d5-32a6ed516ff4 · outbound

This paper cites NAVIG: Natural Language-guided Analysis with Vision Language Models for Image Geo-localization.

Scale, Don't Fine-tune: Guiding Multimodal LLMs for Efficient Visual Place Recognition at Test-Time NAVIG: Natural Language-guided Analysis with Vision Language Models for Image Geo-localization

Reference 26

Resolution
verified exact
local_arxiv, observed 2026-08-05T11:57:17.299591Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-05T11:57:17.245721Z digest=sha256:9dc7e2bb8f72d748a11a52cd7f87a74c923e7a3b0272f0e64ba3b602cce2312c

Observation 3f7f82c3-67a0-4d0c-8d48-1ca3b9fa2d63 · outbound

This paper cites , " * write output.state after.block = add.period write newline.

Scale, Don't Fine-tune: Guiding Multimodal LLMs for Efficient Visual Place Recognition at Test-Time , " * write output.state after.block = add.period write newline

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-05T11:57:17.250154Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T11:57:17.250154Z digest=sha256:9a766dd56afd5284af5c28fb696ccc43d2923230441f4b81173d0b7a67a9b34b

Observation d131aa0b-05ae-48fe-a02e-1a6a066da91a · outbound

This paper cites write newline.

Scale, Don't Fine-tune: Guiding Multimodal LLMs for Efficient Visual Place Recognition at Test-Time write newline

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-05T11:57:17.255005Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T11:57:17.255005Z digest=sha256:eef64147ca676108fdd01f007864ec04eccc53fb76405b20f5794083c60f4cab

Pith citing papers

No inbound Pith citation observations are available.