Pith. sign in

Paper Citation Record · LEDGER

Dynamic Cross-Modal Alignment for Robust Semantic Location Prediction

As of 22 August 2026, this Paper Citation Record lists 14 of 14 outbound references and 0 inbound Pith citation observations for arXiv:2412.09870.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2412.09870 v1

Coverage vector

measured 14 of 14 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-11T16:40:38.970129Z

measured 14 of 14 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-22T06:32:14.747728+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

14 of 14 outbound references displayed

  • verified exact1
  • verified fuzzy3
  • unresolved9
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch1

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 1b89bd27-fe0c-4dec-ab52-b8c76f40118f · outbound

This paper cites Similarity Guided Multimodal Fusion Transformer for Semantic Location Prediction in Social Media.

Dynamic Cross-Modal Alignment for Robust Semantic Location Prediction Similarity Guided Multimodal Fusion Transformer for Semantic Location Prediction in Social Media

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-11T16:40:38.898508Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T16:40:38.898508Z digest=sha256:15cab2c557f770343948ae75f5d40d721fda83942270408946436169b0119b40

Observation 1f0dc90a-fa6e-47f6-9815-d97d33952d16 · outbound

This paper cites Fine-Tuning Large Vision-Language Models as Decision-Making Agents via Reinforcement Learning.

Dynamic Cross-Modal Alignment for Robust Semantic Location Prediction Fine-Tuning Large Vision-Language Models as Decision-Making Agents via Reinforcement Learning

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-11T16:40:38.920563Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T16:40:38.920563Z digest=sha256:b60c6c9bb6d74d7c6bbfbc3b0bf8204e9aababe915896c400674e45881361f7c

Observation 32978308-6454-473b-ae37-2e2278f278ca · outbound

This paper cites An Introduction to Vision-Language Modeling.

Dynamic Cross-Modal Alignment for Robust Semantic Location Prediction An Introduction to Vision-Language Modeling

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-11T16:40:38.926209Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T16:40:38.926209Z digest=sha256:bd63faa72538661c91af54908ed819f8aa9b21eaf402ac40bd4cb89a33212132

Observation 848a0227-cb6e-4b69-bacb-fef31afb94f3 · outbound

This paper cites VisionLLM v2: An End-to-End Generalist Multimodal Large Language Model for Hundreds of Vision-Language Tasks.

Dynamic Cross-Modal Alignment for Robust Semantic Location Prediction VisionLLM v2: An End-to-End Generalist Multimodal Large Language Model for Hundreds of Vision-Language Tasks

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-11T16:40:38.931706Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T16:40:38.931706Z digest=sha256:2db03229cba52eb015624d9d5d03bbf86a82d9c54d155bfd29e791634185fdf8

Observation 521f3395-2251-439e-a251-375270dffeed · outbound

This paper cites TextHawk2: A Large Vision-Language Model Excels in Bilingual OCR and Grounding with 16x Fewer Tokens.

Dynamic Cross-Modal Alignment for Robust Semantic Location Prediction TextHawk2: A Large Vision-Language Model Excels in Bilingual OCR and Grounding with 16x Fewer Tokens

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-11T16:40:38.938311Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T16:40:38.938311Z digest=sha256:5ef342aabfa14735acb2d4a32cbb78229c6bd552e104a157aba76e52bf9c6bb0

Observation 6c262bbb-db2a-4431-a97c-53903ff0b3b8 · outbound

This paper cites Sketch storytelling.

Dynamic Cross-Modal Alignment for Robust Semantic Location Prediction Sketch storytelling

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T16:40:39.472112Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-11T16:40:38.943896Z digest=sha256:3927164d8d9876dcc29721922036fde768bf6fd1c4ff2827866ef310a018690b

Observation 018111f0-e1d5-47d2-a4bb-ab4142c6f7f3 · outbound

This paper cites Towards effective next POI prediction: Spatial and semantic augmentation wit h remote sensing data.

Dynamic Cross-Modal Alignment for Robust Semantic Location Prediction Towards effective next POI prediction: Spatial and semantic augmentation wit h remote sensing data

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T16:40:39.455771Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-11T16:40:38.949121Z digest=sha256:f84818fa7dddc8b97589180db9743cacb7686ce335865823d26d7685d34bc439

Observation 18a14e0b-67a0-45c3-94d7-e1d6c452d6d4 · outbound

This paper cites URL https://doi.org/10.1109/ICDE60146.2024.00104.

Dynamic Cross-Modal Alignment for Robust Semantic Location Prediction URL https://doi.org/10.1109/ICDE60146.2024.00104

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-11T16:40:38.954107Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T16:40:38.954107Z digest=sha256:fab3420982af21ec6686c9775a5f5b348491037e22d953d2b6545afc4ff159c4

Observation 3c345387-06fe-4f8a-8cc0-098bb12088b7 · outbound

This paper cites Large Language Models are Zero-Shot Next Location Predictors.

Dynamic Cross-Modal Alignment for Robust Semantic Location Prediction Large Language Models are Zero-Shot Next Location Predictors

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-11T16:40:38.959118Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T16:40:38.959118Z digest=sha256:d5d097dc2d311bf7948898c0f4269b3ca584e27d4badb52ff83d5f6e0b0073cf

Observation eb4937d5-b114-4936-b046-947acc8b9222 · outbound

This paper cites URL https://doi.org/10.48550/arXiv.2410.09129.

Dynamic Cross-Modal Alignment for Robust Semantic Location Prediction URL https://doi.org/10.48550/arXiv.2410.09129

Reference 13

Resolution
verified exact
doi, observed 2026-08-11T16:40:39.179807Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-11T16:40:38.964907Z digest=sha256:0414b67d44245b5c050a702299e600e37b22b7bef780ac58f5102986ce5fdaa7

Observation 81ef34b3-0edb-4f0d-a3e0-8f599fc03550 · outbound

This paper cites Y ucheng Zhou and Guodong Long.

Dynamic Cross-Modal Alignment for Robust Semantic Location Prediction Y ucheng Zhou and Guodong Long

Reference 2021

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T16:40:39.488706Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-11T16:40:38.909789Z digest=sha256:eff5b38a22323a17b0536c7b88808ba512a4075447a8fffe31008f22a4394951

Observation 5a7b1f18-ff53-4682-aed8-2f50cb567739 · outbound

This paper cites Fine-Tuning Large Vision-Language Models as Decision-Making Agents via Reinforcement Learning.

Dynamic Cross-Modal Alignment for Robust Semantic Location Prediction Fine-Tuning Large Vision-Language Models as Decision-Making Agents via Reinforcement Learning

Reference 2022

Resolution
unresolved
no resolver link, observed 2026-08-11T16:40:38.914836Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T16:40:38.914836Z digest=sha256:a21f92a65271735d1f4b6de2ccb9147b650af7b45d151d06e02cbf0c3985fe11

Observation 99590150-1f20-498d-b55c-7ecd955fdce6 · outbound

This paper cites Thread of Thought Unraveling Chaotic Contexts.

Dynamic Cross-Modal Alignment for Robust Semantic Location Prediction Thread of Thought Unraveling Chaotic Contexts

Reference 2023

Resolution
unresolved
no resolver link, observed 2026-08-11T16:40:38.970129Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T16:40:38.970129Z digest=sha256:25d2b3da1e5e3a604fab5bca6f11a1989de97964359c4b446f8974275ba93b62

Observation 3acd0ae7-bca7-4795-8da1-e51627d00a11 · outbound

This paper cites Similarity Guided Multimodal Fusion Transformer for Semantic Location Prediction in Social Media.

Dynamic Cross-Modal Alignment for Robust Semantic Location Prediction Similarity Guided Multimodal Fusion Transformer for Semantic Location Prediction in Social Media

Reference 2024

Resolution
metadata mismatch
local_arxiv, observed 2026-08-11T16:40:39.305159Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-11T16:40:38.904408Z digest=sha256:1fc4a4c24b8d766f784fa8d954ca42eca39d3c20f69d03d9b4043995297811f7

Pith citing papers

No inbound Pith citation observations are available.