Pith. sign in

Paper Citation Record · LEDGER

Scale, Don't Fine-tune: Guiding Multimodal LLMs for Efficient Visual Place Recognition at Test-Time

As of 6 August 2026, this Paper Citation Record lists 28 of 28 outbound references and 0 inbound Pith citation observations for arXiv:2509.02129.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2509.02129 v1

Coverage vector

measured 28 of 28 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-05T11:57:17.255005Z

measured 28 of 28 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-06T06:34:29.942622+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

28 of 28 outbound references displayed

  • verified exact2
  • verified fuzzy0
  • unresolved26
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation bf1d124e-84b7-41ae-8bdc-4010ad07f9d6 · outbound

This paper cites an unresolved cited work.

Scale, Don't Fine-tune: Guiding Multimodal LLMs for Efficient Visual Place Recognition at Test-Time Unresolved cited work

Reference 1

Resolution
unresolved
raw_fallback, observed 2026-08-05T11:57:17.877913Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-08-05T11:57:16.316998Z digest=sha256:53860218e2fb69a9bc04c1cf58097dc6c6e5c35af8c8e528470a33940a3cea0b

Observation f5110ec7-93fd-4714-b01c-f7e3064378ef · outbound

This paper cites an unresolved cited work.

Scale, Don't Fine-tune: Guiding Multimodal LLMs for Efficient Visual Place Recognition at Test-Time Unresolved cited work

Reference 2

Resolution
unresolved
raw_fallback, observed 2026-08-05T11:57:17.864379Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-08-05T11:57:16.355902Z digest=sha256:9f8dfe9cbffa0c42e74308ccfb9ee5f067d88458559b52c9df337d366a20fb82

Observation fd67a794-0360-486f-a8e0-5680d2dda27f · outbound

This paper cites an unresolved cited work.

Scale, Don't Fine-tune: Guiding Multimodal LLMs for Efficient Visual Place Recognition at Test-Time Unresolved cited work

Reference 3

Resolution
unresolved
raw_fallback, observed 2026-08-05T11:57:17.850435Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-08-05T11:57:16.455330Z digest=sha256:f7618c309ba59e132ce13009e62a742deabe8848e00a3ed3cbbd3e512f63b416

Observation 5990be6f-6b82-403f-a1be-9c77f2b4f35c · outbound

This paper cites an unresolved cited work.

Scale, Don't Fine-tune: Guiding Multimodal LLMs for Efficient Visual Place Recognition at Test-Time Unresolved cited work

Reference 4

Resolution
unresolved
raw_fallback, observed 2026-08-05T11:57:17.836818Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-08-05T11:57:16.519091Z digest=sha256:41d3ba883c1db2c4eb5d5aa47a0b5b6bb5f6f5f94e5b7ef2a1cf518d36e7abf3

Observation e029c155-342d-4b6c-a22b-0cc2dc634b09 · outbound

This paper cites an unresolved cited work.

Scale, Don't Fine-tune: Guiding Multimodal LLMs for Efficient Visual Place Recognition at Test-Time Unresolved cited work

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-05T11:57:16.596950Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T11:57:16.596950Z digest=sha256:ec57d426d9881945cca65890788e66fa46924543f4567b2aae757de4c55a07fc

Observation 87999208-3bcf-46a6-90d6-3171d90122ce · outbound

This paper cites an unresolved cited work.

Scale, Don't Fine-tune: Guiding Multimodal LLMs for Efficient Visual Place Recognition at Test-Time Unresolved cited work

Reference 6

Resolution
unresolved
raw_fallback, observed 2026-08-05T11:57:17.822687Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-08-05T11:57:16.683160Z digest=sha256:e3154df8a8704dcedb599d163d513237ce6363dc5f01f44ec46a76b5e0971386

Observation 14285457-87e1-4d05-a671-243df4b6e228 · outbound

This paper cites Alphazero-like Tree-Search can Guide Large Language Model Decoding and Training.

Scale, Don't Fine-tune: Guiding Multimodal LLMs for Efficient Visual Place Recognition at Test-Time Alphazero-like Tree-Search can Guide Large Language Model Decoding and Training

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-05T11:57:16.716582Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T11:57:16.716582Z digest=sha256:aed77016cdd99ede2f9980ac0cd2ee7e78dd8d462c7fe2300e9ca019d98923b4

Observation 0cff251c-719f-4a14-9b30-ce93e8ae5b13 · outbound

This paper cites an unresolved cited work.

Scale, Don't Fine-tune: Guiding Multimodal LLMs for Efficient Visual Place Recognition at Test-Time Unresolved cited work

Reference 8

Resolution
unresolved
raw_fallback, observed 2026-08-05T11:57:17.808279Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-08-05T11:57:16.800663Z digest=sha256:c5727ea899c2fbb5e048e569343eb9bcb67df4a4c078ba7c1b0b499caf826005

Observation bcf3f1b3-a73c-42a9-be4c-295b3440c589 · outbound

This paper cites an unresolved cited work.

Scale, Don't Fine-tune: Guiding Multimodal LLMs for Efficient Visual Place Recognition at Test-Time Unresolved cited work

Reference 9

Resolution
unresolved
raw_fallback, observed 2026-08-05T11:57:17.794198Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-08-05T11:57:16.955019Z digest=sha256:222a0f44bc368a4e24894e1e15376293ec2302dc5fe40af3b9a3ea76eba6866b

Observation 2b2e8ff4-0239-490e-9170-fb08871dc77f · outbound

This paper cites an unresolved cited work.

Scale, Don't Fine-tune: Guiding Multimodal LLMs for Efficient Visual Place Recognition at Test-Time Unresolved cited work

Reference 10

Resolution
unresolved
raw_fallback, observed 2026-08-05T11:57:17.779771Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-08-05T11:57:17.116224Z digest=sha256:95b2457c1c22924d904772e18931bd0db782ba66082452ac40418bef6634f124

Observation a17a134d-2444-482f-89be-826632a77ee2 · outbound

This paper cites an unresolved cited work.

Scale, Don't Fine-tune: Guiding Multimodal LLMs for Efficient Visual Place Recognition at Test-Time Unresolved cited work

Reference 11

Resolution
unresolved
raw_fallback, observed 2026-08-05T11:57:17.764029Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-08-05T11:57:17.120543Z digest=sha256:c0ab5137d6431c21f7fd78c7ba6e124f72374fb925d043ba0e531e8cf61f0187

Observation 6b92c121-f72f-463a-919a-55f1ef69da0f · outbound

This paper cites an unresolved cited work.

Scale, Don't Fine-tune: Guiding Multimodal LLMs for Efficient Visual Place Recognition at Test-Time Unresolved cited work

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-05T11:57:17.185085Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T11:57:17.185085Z digest=sha256:806bd52f6c79fe4e8f049a965185462b5b9d65b3b80deab485905a476fa4f985

Observation 30b25eca-efca-4649-a6c8-d395812f596d · outbound

This paper cites Tell Me Where You Are: Multimodal LLMs Meet Place Recognition.

Scale, Don't Fine-tune: Guiding Multimodal LLMs for Efficient Visual Place Recognition at Test-Time Tell Me Where You Are: Multimodal LLMs Meet Place Recognition

Reference 13

Resolution
verified exact
local_arxiv, observed 2026-08-05T11:57:17.435287Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-08-05T11:57:17.189338Z digest=sha256:cffc4f2d8d0b8d2f351eb971bc4a440977324cf34fc41cd4d60d81aeb1ec8681

Observation e23b4d26-4498-4661-bff8-d1c5facd7bdb · outbound

This paper cites an unresolved cited work.

Scale, Don't Fine-tune: Guiding Multimodal LLMs for Efficient Visual Place Recognition at Test-Time Unresolved cited work

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-05T11:57:17.194612Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T11:57:17.194612Z digest=sha256:dc8919f2e40114d25f4b710d76f0f941db305ecabf29fbd77ab191a401811f9b

Observation c266c17b-4bbe-4c14-a8b9-b450801b6029 · outbound

This paper cites TreeBoN: Enhancing Inference-Time Alignment with Speculative Tree-Search and Best-of-N Sampling.

Scale, Don't Fine-tune: Guiding Multimodal LLMs for Efficient Visual Place Recognition at Test-Time TreeBoN: Enhancing Inference-Time Alignment with Speculative Tree-Search and Best-of-N Sampling

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-05T11:57:17.198660Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T11:57:17.198660Z digest=sha256:a73e6860078fcea9892ff586430bf345f443bbbbe07d4435e90ae63bf1076835

Observation 509ee550-866c-4896-8ebe-960f1fe8c9d4 · outbound

This paper cites Scaling LLM Test-Time Compute Optimally can be More Effective than Scaling Model Parameters.

Scale, Don't Fine-tune: Guiding Multimodal LLMs for Efficient Visual Place Recognition at Test-Time Scaling LLM Test-Time Compute Optimally can be More Effective than Scaling Model Parameters

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-05T11:57:17.203309Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T11:57:17.203309Z digest=sha256:b199ef2072c607c596335bf17f276e2bd4474b925c3113983a07d5e599f79d46

Observation 04a8acd0-7b26-4efa-854d-20f1bf1a13b3 · outbound

This paper cites an unresolved cited work.

Scale, Don't Fine-tune: Guiding Multimodal LLMs for Efficient Visual Place Recognition at Test-Time Unresolved cited work

Reference 17

Resolution
unresolved
raw_fallback, observed 2026-08-05T11:57:17.749532Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-08-05T11:57:17.208210Z digest=sha256:6f35b1dd8b5c0ea92f1b47efcf6be0c516629f2f48e66e32bd7861d8982ab180

Observation f9a92daa-1f23-4d7f-ac87-7567485d0efa · outbound

This paper cites an unresolved cited work.

Scale, Don't Fine-tune: Guiding Multimodal LLMs for Efficient Visual Place Recognition at Test-Time Unresolved cited work

Reference 18

Resolution
unresolved
raw_fallback, observed 2026-08-05T11:57:17.734489Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-08-05T11:57:17.212218Z digest=sha256:1455cadf11aef444ea51806d5513582ceebb177e5dac9c50ed2b7e7a73537eed

Observation 65273ceb-6a4d-4eb7-ab27-4ff2ade810b1 · outbound

This paper cites an unresolved cited work.

Scale, Don't Fine-tune: Guiding Multimodal LLMs for Efficient Visual Place Recognition at Test-Time Unresolved cited work

Reference 19

Resolution
unresolved
raw_fallback, observed 2026-08-05T11:57:17.720097Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-08-05T11:57:17.216064Z digest=sha256:470507450ffa7ce6a0f975997c5ef4dcecaf33f280b5ff73c174acb0c53fce1f

Observation 324a6c31-86b4-420a-888f-0b02322b2b8e · outbound

This paper cites an unresolved cited work.

Scale, Don't Fine-tune: Guiding Multimodal LLMs for Efficient Visual Place Recognition at Test-Time Unresolved cited work

Reference 20

Resolution
unresolved
raw_fallback, observed 2026-08-05T11:57:17.705310Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-08-05T11:57:17.220188Z digest=sha256:7a153ee4e2ad4e6e1d1628c504dac30b4110f7f7c1c1e1c30be59470a8504cc1

Observation 1d75041c-d04c-407a-a219-1a13e8945aee · outbound

This paper cites an unresolved cited work.

Scale, Don't Fine-tune: Guiding Multimodal LLMs for Efficient Visual Place Recognition at Test-Time Unresolved cited work

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-05T11:57:17.224316Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T11:57:17.224316Z digest=sha256:70369399fe5630f21468a412e27355fe7ed0301936d051548db8d5dde2a900b1

Observation 60d893ea-1597-4763-9791-91d1b4f62f40 · outbound

This paper cites an unresolved cited work.

Scale, Don't Fine-tune: Guiding Multimodal LLMs for Efficient Visual Place Recognition at Test-Time Unresolved cited work

Reference 22

Resolution
unresolved
raw_fallback, observed 2026-08-05T11:57:17.680668Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-08-05T11:57:17.228965Z digest=sha256:d807d43c32a8bef7db92bb5cbbf2b789e456e071326a2260990038695a87b81d

Observation 40a982e5-ef36-4df9-9b57-8794f38db75d · outbound

This paper cites an unresolved cited work.

Scale, Don't Fine-tune: Guiding Multimodal LLMs for Efficient Visual Place Recognition at Test-Time Unresolved cited work

Reference 23

Resolution
unresolved
raw_fallback, observed 2026-08-05T11:57:17.666062Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-08-05T11:57:17.232985Z digest=sha256:beacf3bedf21dc8d2e53e95cb7791ad27421fdcf86feca54634b7da06aa3c4b0

Observation 420fffb0-6471-4d2c-9441-8ffda85c88f6 · outbound

This paper cites an unresolved cited work.

Scale, Don't Fine-tune: Guiding Multimodal LLMs for Efficient Visual Place Recognition at Test-Time Unresolved cited work

Reference 24

Resolution
unresolved
raw_fallback, observed 2026-08-05T11:57:17.651568Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-08-05T11:57:17.237103Z digest=sha256:2529bea32249eb51e6f80bf520376fefb66ab68790751ec5c941651bdfd7bd43

Observation 1e339af1-e62b-49c2-8f8b-6aa90c937897 · outbound

This paper cites an unresolved cited work.

Scale, Don't Fine-tune: Guiding Multimodal LLMs for Efficient Visual Place Recognition at Test-Time Unresolved cited work

Reference 25

Resolution
unresolved
raw_fallback, observed 2026-08-05T11:57:17.636106Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-08-05T11:57:17.241091Z digest=sha256:dfbf57e34498e5bdc2727ecd76f0a5eebbf8e8905bc7f83f0f45916baf6e48c0

Observation ee72edad-a407-4f03-a9d5-32a6ed516ff4 · outbound

This paper cites NAVIG: Natural Language-guided Analysis with Vision Language Models for Image Geo-localization.

Scale, Don't Fine-tune: Guiding Multimodal LLMs for Efficient Visual Place Recognition at Test-Time NAVIG: Natural Language-guided Analysis with Vision Language Models for Image Geo-localization

Reference 26

Resolution
verified exact
local_arxiv, observed 2026-08-05T11:57:17.299591Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-08-05T11:57:17.245721Z digest=sha256:d0bf5619a6f7cc2608c525b7e6e26bd259b5c48a4ec9fba3b8aade66c142fab1

Observation 3f7f82c3-67a0-4d0c-8d48-1ca3b9fa2d63 · outbound

This paper cites , " * write output.state after.block = add.period write newline.

Scale, Don't Fine-tune: Guiding Multimodal LLMs for Efficient Visual Place Recognition at Test-Time , " * write output.state after.block = add.period write newline

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-05T11:57:17.250154Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T11:57:17.250154Z digest=sha256:afaea177e343af77df4947dc2a470ddf17846d945e804757f0dcfe41f769dd13

Observation d131aa0b-05ae-48fe-a02e-1a6a066da91a · outbound

This paper cites write newline.

Scale, Don't Fine-tune: Guiding Multimodal LLMs for Efficient Visual Place Recognition at Test-Time write newline

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-05T11:57:17.255005Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T11:57:17.255005Z digest=sha256:0df78e1cf69b73fdfa8e6d5777e7d72e6ca0c2c7a0b631d2e1738e9a6b41ebc9

Pith citing papers

No inbound Pith citation observations are available.