Pith. sign in

Paper Citation Record · LEDGER

Fine-Grained Food Image Understanding via Target-Aware Data Alignment

As of 10 August 2026, this Paper Citation Record lists 22 of 22 outbound references and 0 inbound Pith citation observations for arXiv:2607.25794.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2607.25794 v1

Coverage vector

measured 22 of 22 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-01T01:32:33.691163Z

measured 22 of 22 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-10T06:31:04.303077+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

22 of 22 outbound references displayed

  • verified exact1
  • verified fuzzy0
  • unresolved21
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 71822307-3b9a-4fe5-9b57-976d2447c888 · outbound

This paper cites Large scale visual food recognition,.

Fine-Grained Food Image Understanding via Target-Aware Data Alignment Large scale visual food recognition,

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-01T01:32:30.850883Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T01:32:30.850883Z digest=sha256:8db37d9de150c911e026363b4ca59f3467b9e373ae926d2e90999e065306015e

Observation b909d30d-0a76-4638-bf09-67ba3619bf5c · outbound

This paper cites Food-101: Mining discriminative components with random forests,.

Fine-Grained Food Image Understanding via Target-Aware Data Alignment Food-101: Mining discriminative components with random forests,

Reference 2

Resolution
verified exact
doi, observed 2026-08-01T01:36:19.906506Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-01T01:32:30.928635Z digest=sha256:6560cfe9668265f1152e9edc6fd19fe4223f403f123adbacf6f113d0a0e41a87

Observation 3b0b58e3-1d0a-4c29-8752-b542e742649b · outbound

This paper cites Dishcovery Mission II Challenge: Where VLM meets food,.

Fine-Grained Food Image Understanding via Target-Aware Data Alignment Dishcovery Mission II Challenge: Where VLM meets food,

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-01T01:32:31.018574Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T01:32:31.018574Z digest=sha256:3d33458b76418a3059a83f94aebb598d9c6c6386bc7c121602de7e926c6e4334

Observation 5011870a-4062-45ec-b15f-5a45f7943446 · outbound

This paper cites Learning transferable visual models from natural language supervision,.

Fine-Grained Food Image Understanding via Target-Aware Data Alignment Learning transferable visual models from natural language supervision,

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-01T01:32:31.096810Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T01:32:31.096810Z digest=sha256:4f43ed9f697273835feb757b0f6aa60528de69b483af909cb3da1913345cea6e

Observation 1d4e9d48-164c-41fd-87cb-6cfa0f79bd8c · outbound

This paper cites Scaling up visual and vision-language representation learning with noisy text supervision,.

Fine-Grained Food Image Understanding via Target-Aware Data Alignment Scaling up visual and vision-language representation learning with noisy text supervision,

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-01T01:32:31.209122Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T01:32:31.209122Z digest=sha256:862a5b441cda0f300d7686fc43993c4bb1750b93caa32d0803ecac177ee17d89

Observation ec4cc4b1-1195-4617-88c6-7aede3e7e100 · outbound

This paper cites OpenCLIP,.

Fine-Grained Food Image Understanding via Target-Aware Data Alignment OpenCLIP,

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-01T01:32:31.316247Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T01:32:31.316247Z digest=sha256:c5e4000acd1597019ff14c10d1abe18586e201a3375f905189093f685cc3d7a7

Observation 4bc8373a-9368-4d7a-9109-a9e0fb6563d3 · outbound

This paper cites Sigmoid loss for language image pre-training,.

Fine-Grained Food Image Understanding via Target-Aware Data Alignment Sigmoid loss for language image pre-training,

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-01T01:32:31.413710Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T01:32:31.413710Z digest=sha256:605c2e06d56fd9bb9c352ab9c69c0b6bdb9505dec44dbfe07b8ca1948c8f84c1

Observation 2fb9aa6e-023d-45ea-9f50-5ea57f34cbfe · outbound

This paper cites Recipe1M+: A dataset for learning cross-modal embeddings for cooking recipes and food images,.

Fine-Grained Food Image Understanding via Target-Aware Data Alignment Recipe1M+: A dataset for learning cross-modal embeddings for cooking recipes and food images,

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-01T01:32:31.524269Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T01:32:31.524269Z digest=sha256:3e02bd3b27f6c28f5baf6522d1eb21c4ef3f1c1cff1602480d7495baf0ec1b91

Observation 4113570c-a919-4009-8331-caa84c3b06da · outbound

This paper cites Passage Re-ranking with BERT.

Fine-Grained Food Image Understanding via Target-Aware Data Alignment Passage Re-ranking with BERT

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-01T01:32:31.632993Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T01:32:31.632993Z digest=sha256:39a0a188a9848c86fb01d293239a04256c40a639e4b7cc6caf0ad5e792c86187

Observation 50866322-c439-4de3-ac5b-987a2b194ab6 · outbound

This paper cites Reciprocal rank fusion outperforms Condorcet and individual rank learning methods,.

Fine-Grained Food Image Understanding via Target-Aware Data Alignment Reciprocal rank fusion outperforms Condorcet and individual rank learning methods,

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-01T01:32:31.737098Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T01:32:31.737098Z digest=sha256:c549e0cd5ff17a627a595d4ec4642f1a926692fe2ab3c3675ab7a9068470a037

Observation e08251a9-de2f-4b18-8d70-ade1ef08f85e · outbound

This paper cites Is ChatGPT good at search? Investigating large language models as re-ranking agents,.

Fine-Grained Food Image Understanding via Target-Aware Data Alignment Is ChatGPT good at search? Investigating large language models as re-ranking agents,

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-01T01:32:31.905207Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T01:32:31.905207Z digest=sha256:7300d1d239ec98d0b428a4c06d06e7b6ea2a44b12f62b1e0a2e747c14cb999f0

Observation ee499a87-f617-4c8f-965c-d14822c0ec41 · outbound

This paper cites CLIP-Adapter: Better vision-language models with feature adapters,.

Fine-Grained Food Image Understanding via Target-Aware Data Alignment CLIP-Adapter: Better vision-language models with feature adapters,

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-01T01:32:32.026185Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T01:32:32.026185Z digest=sha256:2ca07f6c188e451518afc2fccd3292e9822d685da69e826e045190e5dd5b5117

Observation a0e35696-986c-463d-9c07-23fc58da0d66 · outbound

This paper cites LoRA: Low-rank adaptation of large language models,.

Fine-Grained Food Image Understanding via Target-Aware Data Alignment LoRA: Low-rank adaptation of large language models,

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-01T01:32:32.233068Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T01:32:32.233068Z digest=sha256:c0c82250c8e6e9be1f7f33186eefa8d754429b24a41a604e114bf4ee7765e919

Observation c9e4422c-3f6b-41b8-921f-b87e4ef6b8dc · outbound

This paper cites DoRA: Weight-decomposed low-rank adaptation,.

Fine-Grained Food Image Understanding via Target-Aware Data Alignment DoRA: Weight-decomposed low-rank adaptation,

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-01T01:32:32.353401Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T01:32:32.353401Z digest=sha256:debab38d7e28eb73940283225ab8952a31adf7d1819f137199dc458d1ff9f938

Observation 8daaf14e-c364-4fc4-8826-61df1af6d98f · outbound

This paper cites When and why vision-language models behave like bags-of-words, and what to do about it?.

Fine-Grained Food Image Understanding via Target-Aware Data Alignment When and why vision-language models behave like bags-of-words, and what to do about it?

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-01T01:32:32.509324Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T01:32:32.509324Z digest=sha256:6d3902ddefb28e40a19904c98bc768207e0fe5603d25b70b4db4d6f639195565

Observation 24415aea-59b0-4d02-a711-2fcf757b1fb2 · outbound

This paper cites Robust fine-tuning of zero-shot models,.

Fine-Grained Food Image Understanding via Target-Aware Data Alignment Robust fine-tuning of zero-shot models,

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-01T01:32:32.632467Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T01:32:32.632467Z digest=sha256:dca9b37a8379b215ee20a667f790676dd711509069e8086808461445dd06aed5

Observation d39d5fe0-49df-43cb-b1a7-e0731b559b5b · outbound

This paper cites Data filtering networks,.

Fine-Grained Food Image Understanding via Target-Aware Data Alignment Data filtering networks,

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-01T01:32:32.802189Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T01:32:32.802189Z digest=sha256:5829f93e4758b867bded32f9382334b7dfdea5d353b06afa67e9d803da9856a0

Observation 6d0e6292-82af-4548-95f5-d21db5e3b9d6 · outbound

This paper cites Meta CLIP 2: A Worldwide Scaling Recipe.

Fine-Grained Food Image Understanding via Target-Aware Data Alignment Meta CLIP 2: A Worldwide Scaling Recipe

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-01T01:32:32.988872Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T01:32:32.988872Z digest=sha256:80a421fa307556f912162bae5fcd9a88430b0208f43b08d88b3e5c595bf6f750

Observation 757bf3f9-2c13-4fcc-ab7b-661a22738357 · outbound

This paper cites Gemma 4 31B Instruct,.

Fine-Grained Food Image Understanding via Target-Aware Data Alignment Gemma 4 31B Instruct,

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-01T01:32:33.172208Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T01:32:33.172208Z digest=sha256:a434b442c46f54be04bc3162fea0f559349f4b04e751e3e62292ccf803186ab9

Observation e028e063-5105-46da-8d61-23c2e401d404 · outbound

This paper cites Representation Learning with Contrastive Predictive Coding.

Fine-Grained Food Image Understanding via Target-Aware Data Alignment Representation Learning with Contrastive Predictive Coding

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-01T01:32:33.376887Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T01:32:33.376887Z digest=sha256:3b41caea95b2adf3cb609c0d3600b3ff69f90c56343c3075afa6dfa57e4cf53e

Observation 1aea730e-f5c5-4f3f-be6b-ca7ceb03c6f2 · outbound

This paper cites Hugging Face Hub,.

Fine-Grained Food Image Understanding via Target-Aware Data Alignment Hugging Face Hub,

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-01T01:32:33.691163Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T01:32:33.691163Z digest=sha256:6a8bc5435a20e0a0a74bca82f429110e0491859cebeb9cd662b959e2154fc0a7

Observation 769b9aa2-e3dd-4831-ba80-cba6f28b0042 · outbound

This paper cites Representation Learning with Contrastive Predictive Coding.

Fine-Grained Food Image Understanding via Target-Aware Data Alignment Representation Learning with Contrastive Predictive Coding

Reference 2018

Resolution
unresolved
no resolver link, observed 2026-08-01T01:32:33.548750Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T01:32:33.548750Z digest=sha256:bf77c28a834f0d0d03e571cdf201bd972ee942f1270d0e6140235bfb2e6dcec7

Pith citing papers

No inbound Pith citation observations are available.