Pith. sign in

Paper Citation Record · LEDGER

Dynamic Cross-Modal Alignment for Robust Semantic Location Prediction

As of 16 August 2026, this Paper Citation Record lists 14 of 14 outbound references and 0 inbound Pith citation observations for arXiv:2412.09870.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2412.09870 v1

Coverage vector

measured 14 of 14 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-11T16:40:38.970129Z

measured 14 of 14 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-16T06:30:59.297886+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

14 of 14 outbound references displayed

  • verified exact1
  • verified fuzzy3
  • unresolved9
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch1

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 1b89bd27-fe0c-4dec-ab52-b8c76f40118f · outbound

This paper cites Similarity Guided Multimodal Fusion Transformer for Semantic Location Prediction in Social Media.

Dynamic Cross-Modal Alignment for Robust Semantic Location Prediction Similarity Guided Multimodal Fusion Transformer for Semantic Location Prediction in Social Media

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-11T16:40:38.898508Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T16:40:38.898508Z digest=sha256:ea7a3c827bfd1c4a424413ba72f2b9dc520b2c349701e2f394e0828e5bb37185

Observation 1f0dc90a-fa6e-47f6-9815-d97d33952d16 · outbound

This paper cites Fine-Tuning Large Vision-Language Models as Decision-Making Agents via Reinforcement Learning.

Dynamic Cross-Modal Alignment for Robust Semantic Location Prediction Fine-Tuning Large Vision-Language Models as Decision-Making Agents via Reinforcement Learning

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-11T16:40:38.920563Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T16:40:38.920563Z digest=sha256:d74131becf2762590b7ebae1d643120bf042a66504320254739decb1f488b9a7

Observation 32978308-6454-473b-ae37-2e2278f278ca · outbound

This paper cites An Introduction to Vision-Language Modeling.

Dynamic Cross-Modal Alignment for Robust Semantic Location Prediction An Introduction to Vision-Language Modeling

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-11T16:40:38.926209Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T16:40:38.926209Z digest=sha256:7f4cf95fc20a77b5cfbdbd8d2d6fb4c81995fdb517cd6abc326f7627cc3bf213

Observation 848a0227-cb6e-4b69-bacb-fef31afb94f3 · outbound

This paper cites VisionLLM v2: An End-to-End Generalist Multimodal Large Language Model for Hundreds of Vision-Language Tasks.

Dynamic Cross-Modal Alignment for Robust Semantic Location Prediction VisionLLM v2: An End-to-End Generalist Multimodal Large Language Model for Hundreds of Vision-Language Tasks

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-11T16:40:38.931706Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T16:40:38.931706Z digest=sha256:84096b4e3385d7f24eedd4ed55ab7736e2fb250d7c15c85efc4bd9080ac9d018

Observation 521f3395-2251-439e-a251-375270dffeed · outbound

This paper cites TextHawk2: A Large Vision-Language Model Excels in Bilingual OCR and Grounding with 16x Fewer Tokens.

Dynamic Cross-Modal Alignment for Robust Semantic Location Prediction TextHawk2: A Large Vision-Language Model Excels in Bilingual OCR and Grounding with 16x Fewer Tokens

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-11T16:40:38.938311Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T16:40:38.938311Z digest=sha256:df62045488d63a51561c6652cc9dbc93677f6b52be57fbf7c451a8f2d5b34348

Observation 6c262bbb-db2a-4431-a97c-53903ff0b3b8 · outbound

This paper cites Sketch storytelling.

Dynamic Cross-Modal Alignment for Robust Semantic Location Prediction Sketch storytelling

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T16:40:39.472112Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-11T16:40:38.943896Z digest=sha256:577c2c7dfbf4751d9d9bf4c472fead27d1e20b5716db047fcc89bc11d7740f69

Observation 018111f0-e1d5-47d2-a4bb-ab4142c6f7f3 · outbound

This paper cites Towards effective next POI prediction: Spatial and semantic augmentation wit h remote sensing data.

Dynamic Cross-Modal Alignment for Robust Semantic Location Prediction Towards effective next POI prediction: Spatial and semantic augmentation wit h remote sensing data

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T16:40:39.455771Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-11T16:40:38.949121Z digest=sha256:6f422ade0dce7dfed21e5ba122749a8ea635805234d5b579cf4b74c7775831a1

Observation 18a14e0b-67a0-45c3-94d7-e1d6c452d6d4 · outbound

This paper cites URL https://doi.org/10.1109/ICDE60146.2024.00104.

Dynamic Cross-Modal Alignment for Robust Semantic Location Prediction URL https://doi.org/10.1109/ICDE60146.2024.00104

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-11T16:40:38.954107Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T16:40:38.954107Z digest=sha256:fb21959a6a1fc1043b701e53f2eed33be75c4ecc434d51d2a00c2c4ebc11cf11

Observation 3c345387-06fe-4f8a-8cc0-098bb12088b7 · outbound

This paper cites Large Language Models are Zero-Shot Next Location Predictors.

Dynamic Cross-Modal Alignment for Robust Semantic Location Prediction Large Language Models are Zero-Shot Next Location Predictors

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-11T16:40:38.959118Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T16:40:38.959118Z digest=sha256:75794996ef09e974cd8a3e7fca7800ec0c39af8be9e6ad42abed2954e5d3c22b

Observation eb4937d5-b114-4936-b046-947acc8b9222 · outbound

This paper cites URL https://doi.org/10.48550/arXiv.2410.09129.

Dynamic Cross-Modal Alignment for Robust Semantic Location Prediction URL https://doi.org/10.48550/arXiv.2410.09129

Reference 13

Resolution
verified exact
doi, observed 2026-08-11T16:40:39.179807Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-11T16:40:38.964907Z digest=sha256:cf5843cee8a5b242b697d04431be351cc36d54c489ae61ec237e9ecea51dc77c

Observation 81ef34b3-0edb-4f0d-a3e0-8f599fc03550 · outbound

This paper cites Y ucheng Zhou and Guodong Long.

Dynamic Cross-Modal Alignment for Robust Semantic Location Prediction Y ucheng Zhou and Guodong Long

Reference 2021

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T16:40:39.488706Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-11T16:40:38.909789Z digest=sha256:30f472c58a0e9f1a27a854b87b305b92cb85f4f33db2e03f947c53bbeec0e718

Observation 5a7b1f18-ff53-4682-aed8-2f50cb567739 · outbound

This paper cites Fine-Tuning Large Vision-Language Models as Decision-Making Agents via Reinforcement Learning.

Dynamic Cross-Modal Alignment for Robust Semantic Location Prediction Fine-Tuning Large Vision-Language Models as Decision-Making Agents via Reinforcement Learning

Reference 2022

Resolution
unresolved
no resolver link, observed 2026-08-11T16:40:38.914836Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T16:40:38.914836Z digest=sha256:7f4d9b64a3aa0acd0d45b6d9f3aea81f857f752309f619222e4370db0d01dc53

Observation 99590150-1f20-498d-b55c-7ecd955fdce6 · outbound

This paper cites Thread of Thought Unraveling Chaotic Contexts.

Dynamic Cross-Modal Alignment for Robust Semantic Location Prediction Thread of Thought Unraveling Chaotic Contexts

Reference 2023

Resolution
unresolved
no resolver link, observed 2026-08-11T16:40:38.970129Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T16:40:38.970129Z digest=sha256:2d8e0f6793c7428748019a944e18308b0ff1de8ab182c583077c54df2b2de4e0

Observation 3acd0ae7-bca7-4795-8da1-e51627d00a11 · outbound

This paper cites Similarity Guided Multimodal Fusion Transformer for Semantic Location Prediction in Social Media.

Dynamic Cross-Modal Alignment for Robust Semantic Location Prediction Similarity Guided Multimodal Fusion Transformer for Semantic Location Prediction in Social Media

Reference 2024

Resolution
metadata mismatch
local_arxiv, observed 2026-08-11T16:40:39.305159Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-11T16:40:38.904408Z digest=sha256:5bd76983d038563d9faffb3132692d829197aec3f011c6cca7e4ba257c9fea5a

Pith citing papers

No inbound Pith citation observations are available.