Pith. sign in

Paper Citation Record · LEDGER

Evaluating Recabilities of Foundation Models: A Multi-Domain, Multi-Dataset Benchmark

As of 6 August 2026, this Paper Citation Record lists 25 of 25 outbound references and 1 inbound Pith citation observation for arXiv:2508.21354.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2508.21354 v1

Coverage vector

measured 25 of 25 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-05T14:24:24.931453Z

measured 26 of 26 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-06T06:34:29.942622+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-05T11:13:40.751517Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-08-05T11:13:44.946863Z

Reference resolution

25 of 25 outbound references displayed

  • verified exact4
  • verified fuzzy6
  • unresolved13
  • parse uncertain0
  • malformed identifier1
  • metadata mismatch1

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation cd24d7a3-d9bd-47c5-a2a5-2d5fe33d308c · outbound

This paper cites DeepSeek LLM: Scaling Open-Source Language Models with Longtermism.

Evaluating Recabilities of Foundation Models: A Multi-Domain, Multi-Dataset Benchmark DeepSeek LLM: Scaling Open-Source Language Models with Longtermism

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-05T14:24:22.828300Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T14:24:22.828300Z digest=sha256:852ddcb2a00a093907eacfe9ab7a59d4299de0e2761805c04ef4d85adb78710e

Observation 4ffbc0a2-0e72-47ba-bb8b-f99f0203fbcf · outbound

This paper cites Revisiting Word Embeddings in the LLM Era.

Evaluating Recabilities of Foundation Models: A Multi-Domain, Multi-Dataset Benchmark Revisiting Word Embeddings in the LLM Era

Reference 4

Resolution
verified exact
local_arxiv, observed 2026-08-05T14:24:26.533314Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-05T14:24:23.159373Z digest=sha256:c3be6d58a0704a359c37a7dc4686025b9ebc42b05ddb9dc60131834c14846390

Observation 296096be-3a16-4b79-840c-d3833d029893 · outbound

This paper cites Cross-domain recom- mendation via cluster-level latent factor model.

Evaluating Recabilities of Foundation Models: A Multi-Domain, Multi-Dataset Benchmark Cross-domain recom- mendation via cluster-level latent factor model

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T14:24:27.565748Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-05T14:24:23.275635Z digest=sha256:f7097c7ff62807eede04468d37b06e8a3873e3f9bbba4e68526d6042ed9ffc1c

Observation 432d37dd-ca8b-43d2-a880-e5b918dbd239 · outbound

This paper cites DA-GCN: A Domain-aware Attentive Graph Convolution Network for Shared-account Cross-domain Sequential Recommendation.

Evaluating Recabilities of Foundation Models: A Multi-Domain, Multi-Dataset Benchmark DA-GCN: A Domain-aware Attentive Graph Convolution Network for Shared-account Cross-domain Sequential Recommendation

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-05T14:24:23.416343Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T14:24:23.416343Z digest=sha256:13fa5e6badcc5794ac7c89f52436b073bd3b8c33392f90f04a7a7455a7cab4ba

Observation f0fe05b8-c40f-4c3b-a1e0-261ca3f49549 · outbound

This paper cites Beyond Utility: Evaluating LLM as Recommender.

Evaluating Recabilities of Foundation Models: A Multi-Domain, Multi-Dataset Benchmark Beyond Utility: Evaluating LLM as Recommender

Reference 9

Resolution
verified exact
local_arxiv, observed 2026-08-05T14:24:26.199987Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-05T14:24:23.572582Z digest=sha256:f1621159fddc596beaeafe2b9ee61b127079579df0913a053dcd23b6b88cabf0

Observation 270f586d-2c70-4437-8379-e1536549d8de · outbound

This paper cites LLaVA-OneVision: Easy Visual Task Transfer.

Evaluating Recabilities of Foundation Models: A Multi-Domain, Multi-Dataset Benchmark LLaVA-OneVision: Easy Visual Task Transfer

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-05T14:24:23.656589Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T14:24:23.656589Z digest=sha256:3e3978015ba961b8838884a163a37d83433000784f5a220032fe210b5168984c

Observation 454176b1-a7b8-4305-8ce7-604688e57b73 · outbound

This paper cites Language model evolutionary algorithms for recommender systems: Benchmarks and algorithm comparisons.

Evaluating Recabilities of Foundation Models: A Multi-Domain, Multi-Dataset Benchmark Language model evolutionary algorithms for recommender systems: Benchmarks and algorithm comparisons

Reference 11

Resolution
verified exact
raw_fallback, observed 2026-08-05T14:24:25.989145Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-05T14:24:23.756481Z digest=sha256:5f60d4a438275870ac3a9ea390d84d2b575dea1f566bc6a4445fdb2e9cda2bcd

Observation 2a9a03a3-f8ef-43f0-a594-fd31154bfae9 · outbound

This paper cites Learning Multi-Aspect Item Palette: A Semantic Tokenization Framework for Generative Recommendation.

Evaluating Recabilities of Foundation Models: A Multi-Domain, Multi-Dataset Benchmark Learning Multi-Aspect Item Palette: A Semantic Tokenization Framework for Generative Recommendation

Reference 12

Resolution
metadata mismatch
local_arxiv, observed 2026-08-05T14:24:25.704029Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-05T14:24:23.831866Z digest=sha256:8ffe28beb18447667752d3c051c8bd584d039678939b7c499d81c6e1efcf9bc5

Observation 9e86968c-dee9-44df-a41b-9302cd6da27a · outbound

This paper cites Hoang Ngo and Dat Quoc Nguyen.

Evaluating Recabilities of Foundation Models: A Multi-Domain, Multi-Dataset Benchmark Hoang Ngo and Dat Quoc Nguyen

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T14:24:27.276982Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-05T14:24:23.884903Z digest=sha256:457753b4c25789aa02eec3fbec315c96d93effa2dbae1654e0a0bebd37ff4b12

Observation 7f9fe0d4-2ca8-44b0-8a74-9eaad73f86f8 · outbound

This paper cites Gemini 1.5: Unlocking multimodal understanding across millions of tokens of context.

Evaluating Recabilities of Foundation Models: A Multi-Domain, Multi-Dataset Benchmark Gemini 1.5: Unlocking multimodal understanding across millions of tokens of context

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-05T14:24:23.949796Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T14:24:23.949796Z digest=sha256:8387987c1266cb92f4b2a8f88f59cbb7aa77f25dea4f6d249b6ea3c5e6674246

Observation 7062efb2-2a0d-41e8-9090-ee26b38ceb14 · outbound

This paper cites LLaMA: Open and Efficient Foundation Language Models.

Evaluating Recabilities of Foundation Models: A Multi-Domain, Multi-Dataset Benchmark LLaMA: Open and Efficient Foundation Language Models

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-05T14:24:24.033316Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T14:24:24.033316Z digest=sha256:aec9a008620c2b1a0c8a13c23836cad836d23b72a1f69e242d75e76cbd7eff48

Observation 43b18f6b-b4c0-4bd1-bf96-58f60edd7fd0 · outbound

This paper cites Rethinking the Evaluation for Conversational Recommendation in the Era of Large Language Models.

Evaluating Recabilities of Foundation Models: A Multi-Domain, Multi-Dataset Benchmark Rethinking the Evaluation for Conversational Recommendation in the Era of Large Language Models

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-05T14:24:24.104785Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T14:24:24.104785Z digest=sha256:d864490a60520ca5a3a5c9412fab4c883ca832ae8e6fa0116d09c75625439c59

Observation c220245a-184c-45d5-8b9e-9a3026985f8d · outbound

This paper cites HuggingFace's Transformers: State-of-the-art Natural Language Processing.

Evaluating Recabilities of Foundation Models: A Multi-Domain, Multi-Dataset Benchmark HuggingFace's Transformers: State-of-the-art Natural Language Processing

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-05T14:24:24.212069Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T14:24:24.212069Z digest=sha256:eba33440663494da87168e549d7b17dddba82df42793cd063dc857be73e87ada

Observation f450759c-f499-4202-95d5-92069a023be1 · outbound

This paper cites A survey on large language models for recommendation.

Evaluating Recabilities of Foundation Models: A Multi-Domain, Multi-Dataset Benchmark A survey on large language models for recommendation

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T14:24:27.054082Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-05T14:24:24.389531Z digest=sha256:4afff560b34a816ecb134f0c948bdfe6b66a6ec059ade2f085a2718e1b15addf

Observation ce3d5b05-4cbc-44e8-bf90-2cef063a2978 · outbound

This paper cites Qwen2 Technical Report.

Evaluating Recabilities of Foundation Models: A Multi-Domain, Multi-Dataset Benchmark Qwen2 Technical Report

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-05T14:24:24.498733Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T14:24:24.498733Z digest=sha256:a3402b71f2f64c171699768be40e18bf755ac1f1ddf10584ce8a52609f0ddbf2

Observation 8d44194c-7aa7-4aea-87a4-2607d83c1327 · outbound

This paper cites OPT: Open Pre-trained Transformer Language Models.

Evaluating Recabilities of Foundation Models: A Multi-Domain, Multi-Dataset Benchmark OPT: Open Pre-trained Transformer Language Models

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-05T14:24:24.573508Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T14:24:24.573508Z digest=sha256:5c5e516faaf388135031cd7b1efaeac7c917594d9191d24f67742457e6cb667b

Observation 0d9488c5-7796-4d58-b948-62ae9841d120 · outbound

This paper cites Language models as recommender systems: Evalua- tions and limitations.

Evaluating Recabilities of Foundation Models: A Multi-Domain, Multi-Dataset Benchmark Language models as recommender systems: Evalua- tions and limitations

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T14:24:26.915899Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-05T14:24:24.669460Z digest=sha256:71dd06938954c942c86ccef9e75426c04608465b64940a1bfc915febdd45bb12

Observation 4ccaed1e-689a-4b7d-a60e-30461996e99f · outbound

This paper cites 2024.3392335.

Evaluating Recabilities of Foundation Models: A Multi-Domain, Multi-Dataset Benchmark 2024.3392335

Reference 23

Resolution
malformed identifier
no resolver link, observed 2026-08-05T14:24:24.745262Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T14:24:24.745262Z digest=sha256:fc1b1dd8c8175eb11566d42c6c278642bec3983fda822e4e9d674a00ae801269

Observation a8439bf3-dacb-4a59-ac3b-b03c0dad99a3 · outbound

This paper cites C.2 Additional Evaluation Metrics In the main text, we report only the AUC metric due to the space constraints.

Evaluating Recabilities of Foundation Models: A Multi-Domain, Multi-Dataset Benchmark C.2 Additional Evaluation Metrics In the main text, we report only the AUC metric due to the space constraints

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T14:24:26.757690Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-05T14:24:24.931453Z digest=sha256:5d88abefb6583b46ee3afdf5560e9b8c86eb6ca5831fda6e892036fe6042826f

Observation 68d16a66-1a19-4ee6-b588-dac23a91fb01 · outbound

This paper cites The Llama 3 Herd of Models.

Evaluating Recabilities of Foundation Models: A Multi-Domain, Multi-Dataset Benchmark The Llama 3 Herd of Models

Reference 2013

Resolution
unresolved
no resolver link, observed 2026-08-05T14:24:23.081553Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T14:24:23.081553Z digest=sha256:535485f7f88ceb91e2fad7f4550cbf3a61f84767d01347fb9bec5d2ac4e4c4e0

Observation 97fffeb3-4241-456c-ab68-8a4ad6d3a7f6 · outbound

This paper cites TF-DCon: Leveraging Large Language Models (LLMs) to Empower Training-Free Dataset Condensation for Content-Based Recommendation.

Evaluating Recabilities of Foundation Models: A Multi-Domain, Multi-Dataset Benchmark TF-DCon: Leveraging Large Language Models (LLMs) to Empower Training-Free Dataset Condensation for Content-Based Recommendation

Reference 2019

Resolution
verified exact
local_arxiv, observed 2026-08-05T14:24:25.361532Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-05T14:24:24.302375Z digest=sha256:bc4df84eb6dcc2c64defd5acd5b9731adbf37d82cabe3c760675e03ae0974f85

Observation 354ee563-3200-4764-b97d-13d036a6ca52 · outbound

This paper cites ChatGLM: A Family of Large Language Models from GLM-130B to GLM-4 All Tools.

Evaluating Recabilities of Foundation Models: A Multi-Domain, Multi-Dataset Benchmark ChatGLM: A Family of Large Language Models from GLM-130B to GLM-4 All Tools

Reference 2022

Resolution
unresolved
no resolver link, observed 2026-08-05T14:24:23.352696Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T14:24:23.352696Z digest=sha256:740de28e05c50e570b6c1553e62f40af37742ce9c6c5aeeb8e7aa0ae8a5cd033

Observation b2974f55-6519-4442-9380-10258ce9dce8 · outbound

This paper cites Mistral 7B.

Evaluating Recabilities of Foundation Models: A Multi-Domain, Multi-Dataset Benchmark Mistral 7B

Reference 2023

Resolution
unresolved
no resolver link, observed 2026-08-05T14:24:23.504111Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T14:24:23.504111Z digest=sha256:e6070e5d8fd156939348dea3923ac89848ecd6982d284e030d9a06e694bf945a

Observation 541190a3-df68-4da1-bc0b-d6ba231267ea · outbound

This paper cites Differential private knowledge transfer for privacy-preserving cross-domain recommendation.

Evaluating Recabilities of Foundation Models: A Multi-Domain, Multi-Dataset Benchmark Differential private knowledge transfer for privacy-preserving cross-domain recommendation

Reference 2024

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T14:24:27.760426Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-05T14:24:22.981848Z digest=sha256:537ef665f444897a4c4f460b9ba4ddc988cd6b889bb6f48f1b28626e0ef64f21

Observation 4e87e01f-a252-41e5-98a9-84c7adacf045 · outbound

This paper cites Cross-Domain Recommendation: Challenges, Progress, and Prospects.

Evaluating Recabilities of Foundation Models: A Multi-Domain, Multi-Dataset Benchmark Cross-Domain Recommendation: Challenges, Progress, and Prospects

Reference 2025

Resolution
unresolved
no resolver link, observed 2026-08-05T14:24:24.858493Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T14:24:24.858493Z digest=sha256:1c78485c14d1ab24bf402c4a30f0295f7e42d1445181f39394f067e606d71d9d

Pith citing papers

Observation 7917368b-6bf0-4154-9537-afc4211d5087 · inbound

RecBase: Generative Foundation Model Pretraining for Zero-Shot Recommendation cites this paper.

RecBase: Generative Foundation Model Pretraining for Zero-Shot Recommendation Evaluating Recabilities of Foundation Models: A Multi-Domain, Multi-Dataset Benchmark

Reference 29

Resolution
verified exact
local_arxiv, observed 2026-08-05T11:13:45.012281Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-08-05T11:13:40.751517Z digest=sha256:66993e93228143a88550a192e86b7d955d43b87a787eede3b446adf68f636de6