Pith. sign in

Paper Citation Record · LEDGER

TrustLLM: Trustworthiness in Large Language Models

As of 4 August 2026, this Paper Citation Record lists 100 of 299 outbound references and 45 inbound Pith citation observations for arXiv:2401.05561.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2401.05561 v6

Coverage vector

measured 100 of 299 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-05-18T11:17:08.108565Z

measured 145 of 145 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-04T06:34:03.388597+00:00

measured 45 of 45 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-04T01:34:38.960505Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-07-04T19:50:11.226678Z

Reference resolution

100 of 299 outbound references displayed

  • verified exact33
  • verified fuzzy61
  • unresolved2
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch4

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation f2b7546e-1b3a-49c5-bd54-f65d529e97db · outbound

This paper cites A toolkit for text extraction and analysis for natural language processing tasks.

TrustLLM: Trustworthiness in Large Language Models A toolkit for text extraction and analysis for natural language processing tasks

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T11:21:19.609733Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-18T11:17:08.108565Z digest=sha256:fd3851af13d67d4017d06025bff345ff7b2c7b19eded9c4489f156e75c50afd1

Observation 6dc87d8c-cf01-4760-b178-ef5fea9cec60 · outbound

This paper cites Natural language processing: State of the art, current trends and challenges.

TrustLLM: Trustworthiness in Large Language Models Natural language processing: State of the art, current trends and challenges

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T11:21:19.622912Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-18T11:17:08.108565Z digest=sha256:c0e19e28d32345c9ccdeb63519751e095b2c0b7def673e57790e337ffd34c6ff

Observation 52e7a414-e6a3-4e0e-9a7a-4549898e1d30 · outbound

This paper cites Wordcraft: story writing with large language models.

TrustLLM: Trustworthiness in Large Language Models Wordcraft: story writing with large language models

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T11:21:19.616357Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-18T11:17:08.108565Z digest=sha256:d217fc02689853fee5f25cfae355d69f2861f399af0512e7242115e1509744db

Observation 0690adfa-1085-4d54-a8dc-e0568a4ab057 · outbound

This paper cites Multilingual machine translation with large language models: Empirical results and analysis.

TrustLLM: Trustworthiness in Large Language Models Multilingual machine translation with large language models: Empirical results and analysis

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T11:21:19.584706Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-18T11:17:08.108565Z digest=sha256:a85c869cac368aa278e11c935d364c1a5bf6f5d801f8e6aade0534c355a09b46

Observation ded87798-abeb-4063-a612-486c2bcc813d · outbound

This paper cites https://blogs.microsoft.com/blog/2023/02/07/ reinventing-search-with-a-new-ai-powered-microsoft-bing-and-edge-your-copilot-for-the-web/.

TrustLLM: Trustworthiness in Large Language Models https://blogs.microsoft.com/blog/2023/02/07/ reinventing-search-with-a-new-ai-powered-microsoft-bing-and-edge-your-copilot-for-the-web/

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T11:21:19.612931Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-18T11:17:08.108565Z digest=sha256:ca38493945e1203603039d7b4aa45a28a5fe564cc054158a37b818b60baa80e8

Observation 93998411-88da-43a2-8503-f90779080302 · outbound

This paper cites https://medium.com/whatnot-engineering/ enhancing-search-using-large-language-models-f9dcb988bdb9.

TrustLLM: Trustworthiness in Large Language Models https://medium.com/whatnot-engineering/ enhancing-search-using-large-language-models-f9dcb988bdb9

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T11:21:19.603185Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-18T11:17:08.108565Z digest=sha256:335b27ddded720c7feee6e05eafe67e056b4d95ad2990d255bdccc7c28d91968

Observation 7d233665-c017-4261-9e9b-e6cfce56cd87 · outbound

This paper cites WebGPT: Browser-assisted question-answering with human feedback.

TrustLLM: Trustworthiness in Large Language Models WebGPT: Browser-assisted question-answering with human feedback

Reference 7

Resolution
verified exact
local_arxiv, observed 2026-05-18T11:17:08.270100Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-18T11:17:08.108565Z digest=sha256:20b436775f11c8e5a5034c45e4e86cb8e5d8625a7c78830173d422022f74e47d

Observation bff1b0b3-b48b-4511-9cca-a2e87fbf6bf6 · outbound

This paper cites https://www.projectpro.io/article/ large-language-model-use-cases-and-applications/887.

TrustLLM: Trustworthiness in Large Language Models https://www.projectpro.io/article/ large-language-model-use-cases-and-applications/887

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T11:21:19.619333Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-18T11:17:08.108565Z digest=sha256:cf4bb67625e533fed621094454b32765acfafc3b65972312654a3c8aa586c31e

Observation fa095434-dcbc-49ba-98c3-4b197d7891a1 · outbound

This paper cites Code Llama: Open Foundation Models for Code.

TrustLLM: Trustworthiness in Large Language Models Code Llama: Open Foundation Models for Code

Reference 9

Resolution
verified exact
local_arxiv, observed 2026-05-18T11:17:08.278668Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-18T11:17:08.108565Z digest=sha256:1de5ac7c1ab62579006e595f66acc41df9180dd359ef3d884b7bf2b71f7ceae6

Observation c669839d-7755-45d7-a266-8cc15be9f727 · outbound

This paper cites Large language models: The future of b2b software.

TrustLLM: Trustworthiness in Large Language Models Large language models: The future of b2b software

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T11:21:19.581472Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-18T11:17:08.108565Z digest=sha256:3b04654cf07f054c16b1155729d0c8c5403893c789c364f86456b850827beb7b

Observation e64d5080-19d8-48da-ac0a-d32862b21a2b · outbound

This paper cites Bloomberggpt: A large language model for finance.

TrustLLM: Trustworthiness in Large Language Models Bloomberggpt: A large language model for finance

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T11:21:19.588583Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-18T11:17:08.108565Z digest=sha256:ac443763e291115c021d009471e40656298bead8a83d7be5a8276be3ecce11d7

Observation 3ac50255-b411-4afb-b338-061655de1b55 · outbound

This paper cites Scientific discovery in the age of artificial intelligence.

TrustLLM: Trustworthiness in Large Language Models Scientific discovery in the age of artificial intelligence

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T11:21:19.599965Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-18T11:17:08.108565Z digest=sha256:d44f4666adf5bfc746be9de757bc660dfbb314f2230f86e7f345250d638d56aa

Observation b765b235-2134-4094-b845-11eee0929722 · outbound

This paper cites Artificial Intelligence for Science in Quantum, Atomistic, and Continuum Systems.

TrustLLM: Trustworthiness in Large Language Models Artificial Intelligence for Science in Quantum, Atomistic, and Continuum Systems

Reference 13

Resolution
metadata mismatch
arxiv_id, observed 2026-05-18T11:17:08.286332Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-18T11:17:08.108565Z digest=sha256:5821bd9cb2b4df385f16ba84e455d00098058d0ca189e6c6727981405fc3902b

Observation adb79616-6524-49ee-bc3e-d6a3e93235fb · outbound

This paper cites The impact of large language models on scientific discovery: a preliminary study using gpt-4.

TrustLLM: Trustworthiness in Large Language Models The impact of large language models on scientific discovery: a preliminary study using gpt-4

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T11:21:19.606740Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-18T11:17:08.108565Z digest=sha256:3f0956a00c71edf6734e96202323fbe7731920dc6d15023fc0b4d639756556be

Observation ae912b18-7b88-4276-b765-dc92f64e82e3 · outbound

This paper cites Pllama: An open-source large language model for plant science.

TrustLLM: Trustworthiness in Large Language Models Pllama: An open-source large language model for plant science

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T11:21:19.596495Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-18T11:17:08.108565Z digest=sha256:fe2db5bd368a32f863d3ba49b7e6af5b528e6539975ba6583ba25ca201f2b4c1

Observation f581c66b-4a58-4cc9-b75e-7d90ba06919b · outbound

This paper cites The future landscape of large language models in medicine.

TrustLLM: Trustworthiness in Large Language Models The future landscape of large language models in medicine

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T11:21:19.629205Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-18T11:17:08.108565Z digest=sha256:09eb055d5f6c7ce85765ad8a92dfe396ab52b1b38a21d99f1f331b763c3c8381

Observation 3f7610a3-606e-4da1-a98f-aa739951b492 · outbound

This paper cites ChiMed-GPT: A Chinese Medical Large Language Model with Full Training Regime and Better Alignment to Human Preferences.

TrustLLM: Trustworthiness in Large Language Models ChiMed-GPT: A Chinese Medical Large Language Model with Full Training Regime and Better Alignment to Human Preferences

Reference 17

Resolution
verified exact
arxiv_id, observed 2026-05-18T11:17:08.295160Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-18T11:17:08.108565Z digest=sha256:fe0e9ced3cb76f95fa0be111e518f361fa3019a13675a24e02f5b2e716ccde0d

Observation 4ba49faa-1555-4be2-aae5-1e60002dcee3 · outbound

This paper cites Alpacare:instruction-tuned large language models for medical application.

TrustLLM: Trustworthiness in Large Language Models Alpacare:instruction-tuned large language models for medical application

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T11:21:19.626113Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-18T11:17:08.108565Z digest=sha256:73db93ca41eea5b7e8db0dc17bc8c312ca010d7ed1f8a4e89877b4b334d7db39

Observation d5784197-ed8c-49c7-85a7-5df31e2b0a27 · outbound

This paper cites Davison, Quanzheng Li, Yong Chen, Hongfang Liu, and Lichao Sun.

TrustLLM: Trustworthiness in Large Language Models Davison, Quanzheng Li, Yong Chen, Hongfang Liu, and Lichao Sun

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T11:21:19.632345Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-18T11:17:08.108565Z digest=sha256:7cedfdedf702de21ffae8a51cc567d0ac4cbff28823f15e77d06f389489b602e

Observation dbaafa24-760e-48e3-bb3e-b9664a7d4d2f · outbound

This paper cites Bianque: Balancing the questioning and suggestion ability of health llms with multi-turn health conversations polished by chatgpt.

TrustLLM: Trustworthiness in Large Language Models Bianque: Balancing the questioning and suggestion ability of health llms with multi-turn health conversations polished by chatgpt

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T11:21:19.591865Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-18T11:17:08.108565Z digest=sha256:ce3392e70c18d4d28ed076cac3ffd0582ea3298edc3fa04e8052c618ec16116e

Observation 1182bdc9-1239-446d-8c22-a4a0457b2504 · outbound

This paper cites HuatuoGPT, towards Taming Language Model to Be a Doctor.

TrustLLM: Trustworthiness in Large Language Models HuatuoGPT, towards Taming Language Model to Be a Doctor

Reference 21

Resolution
metadata mismatch
arxiv_id, observed 2026-05-18T11:17:08.301485Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-18T11:17:08.108565Z digest=sha256:261b003a62d7a1c6de1004bb6ca7aaf834e86fb3b5ebe320208a102c47902f3d

Observation afd62b14-d9b9-4151-8b19-f3a722bc3ae2 · outbound

This paper cites Chatdoctor: A medical chat model fine-tuned on a large language model meta-ai (llama) using medical domain knowledge.

TrustLLM: Trustworthiness in Large Language Models Chatdoctor: A medical chat model fine-tuned on a large language model meta-ai (llama) using medical domain knowledge

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T11:21:19.821543Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-18T11:17:08.108565Z digest=sha256:dda399d8ca4044eb2ea07c1a328125b9d5cf0be47e1211c7888072a887009098

Observation 7fc96c94-95c3-4a3d-9f4a-48d6bc6dd74e · outbound

This paper cites Medicalgpt: Training medical gpt model.

TrustLLM: Trustworthiness in Large Language Models Medicalgpt: Training medical gpt model

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T11:21:19.808683Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-18T11:17:08.108565Z digest=sha256:8d560719f7ec88564daadb19582e12730c31b3928c5a979a9420b22c8671e3a4

Observation 97a23ea4-43ab-4dfc-a692-b8b1e6958692 · outbound

This paper cites A domain-specific next-generation large language model (llm) or chatgpt is required for biomedical engineering and research.

TrustLLM: Trustworthiness in Large Language Models A domain-specific next-generation large language model (llm) or chatgpt is required for biomedical engineering and research

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T11:21:19.791695Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-18T11:17:08.108565Z digest=sha256:9e7b82bc4d7fbb20a4da2921948645a11ea98981c1287c366d2227f4f7ae8952

Observation 90ef08fb-b6ad-4478-8449-9fbdda9b16e8 · outbound

This paper cites Towards Generalist Biomedical AI.

TrustLLM: Trustworthiness in Large Language Models Towards Generalist Biomedical AI

Reference 25

Resolution
verified exact
arxiv_id, observed 2026-05-18T11:17:08.307716Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-18T11:17:08.108565Z digest=sha256:ab1c8ce4cb610f6a008e5c47604a33ef22785a92268621c5004ef1240eb376e6

Observation 3d5de8ee-4f84-458f-8485-9ef2ba298abe · outbound

This paper cites Large language models and political science.

TrustLLM: Trustworthiness in Large Language Models Large language models and political science

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T11:21:19.799024Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-18T11:17:08.108565Z digest=sha256:0574e507ce7c884b44702f267d70dd71a037ea1636fcf4d7192f67ce6be05376

Observation 3748b433-3850-406c-91fa-c24274639eb0 · outbound

This paper cites https://github.com/irlab-sdu/fuzi.mingcha.

TrustLLM: Trustworthiness in Large Language Models https://github.com/irlab-sdu/fuzi.mingcha

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T11:21:19.717641Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-18T11:17:08.108565Z digest=sha256:b639ea7cff1bd7f926d3c8a86e9d099008631246499bebab57958db8cd033bcd

Observation 68037678-94d9-4c60-9292-0fe812f61ed7 · outbound

This paper cites Disc-lawllm: Fine-tuning large language models for intelligent legal services.

TrustLLM: Trustworthiness in Large Language Models Disc-lawllm: Fine-tuning large language models for intelligent legal services

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T11:21:19.818151Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-18T11:17:08.108565Z digest=sha256:a24ff0a40e069b0a7ee038e09c11bc55c607e87772d49e441437e83ed5bad1c8

Observation e2c4e8d3-d884-4da2-ac9b-f4aa09ad1364 · outbound

This paper cites Chawla, Olaf Wiest, and Xiangliang Zhang.

TrustLLM: Trustworthiness in Large Language Models Chawla, Olaf Wiest, and Xiangliang Zhang

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T11:21:19.888149Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-18T11:17:08.108565Z digest=sha256:a9dd4b33f5cbc04947f72e19037310fdcd4e519a4aeb25c9d4e9cec5d957187a

Observation d3e938d5-ae90-4683-84d2-77f53991f4b5 · outbound

This paper cites Structured Chemistry Reasoning with Large Language Models.

TrustLLM: Trustworthiness in Large Language Models Structured Chemistry Reasoning with Large Language Models

Reference 30

Resolution
verified exact
arxiv_id, observed 2026-05-18T11:17:08.314110Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-18T11:17:08.108565Z digest=sha256:cb388a46a8c02a72022a1b788f926909f4a56d6d2e78ea328b4bb7230d53f579

Observation f8e95e44-d13b-45af-84e2-96bddd224863 · outbound

This paper cites Marinegpt: Unlocking secrets of "ocean" to the public.

TrustLLM: Trustworthiness in Large Language Models Marinegpt: Unlocking secrets of "ocean" to the public

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T11:21:19.844162Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-18T11:17:08.108565Z digest=sha256:ba5a33f76b209bfa2be86381a2e84e7138a3e154e49c33414b7b428efbe222bc

Observation 45553da3-af6d-4af0-b58f-2a5548b9096d · outbound

This paper cites Oceangpt: A large language model for ocean science tasks.

TrustLLM: Trustworthiness in Large Language Models Oceangpt: A large language model for ocean science tasks

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T11:21:19.761402Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-18T11:17:08.108565Z digest=sha256:2782b75c5e916dc6d4b88f8d2b2a7835cd5a389857a4e5a77300e85c59a69e02

Observation c0f7efa4-8aff-444d-9dbe-6ab149e579df · outbound

This paper cites Taoli llama.

TrustLLM: Trustworthiness in Large Language Models Taoli llama

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T11:21:20.018225Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-18T11:17:08.108565Z digest=sha256:d81ae9a6383e08d228eaf546860015f8f43557905db138b7e149279e16f59f84

Observation 9eb43874-decd-4dd4-a38b-96ca1dd365dd · outbound

This paper cites Artgpt-4: Artistic vision-language understanding with adapter-enhanced minigpt-4.

TrustLLM: Trustworthiness in Large Language Models Artgpt-4: Artistic vision-language understanding with adapter-enhanced minigpt-4

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T11:21:19.692779Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-18T11:17:08.108565Z digest=sha256:38f5e0160dba4b5d2b280f7e443c00739abccbea3d42fa4a9b6318d84221c1f6

Observation d62334ea-3b76-4b30-9de9-17dfa9da04d2 · outbound

This paper cites an unresolved cited work.

TrustLLM: Trustworthiness in Large Language Models Unresolved cited work

Reference 35

Resolution
unresolved
raw_fallback, observed 2026-05-18T11:21:19.757073Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-18T11:17:08.108565Z digest=sha256:85c35598c1a30d685e5fa7fa8ff015607684a13ba638489b079d8f7820aa6ee2

Observation 50912041-3ad1-4a01-a88d-59b30c4b3487 · outbound

This paper cites Dai, Orhan Firat, Melvin Johnson, Dmitry Lepikhin, Alexandre Passos, Sia- mak Shakeri, Emanuel Taropa, Paige Bailey, Zhifeng Chen, Eric Chu, Jonathan H.

TrustLLM: Trustworthiness in Large Language Models Dai, Orhan Firat, Melvin Johnson, Dmitry Lepikhin, Alexandre Passos, Sia- mak Shakeri, Emanuel Taropa, Paige Bailey, Zhifeng Chen, Eric Chu, Jonathan H

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T11:21:19.802380Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-18T11:17:08.108565Z digest=sha256:baa08147fbca8ca57b51c3ce2d5d0cd49222820a0b287a8f72c87b14b7a0fd79

Observation 1341a0b8-52d8-4a04-b3ee-852076565bf4 · outbound

This paper cites Palm: Efficiently training massive language models.

TrustLLM: Trustworthiness in Large Language Models Palm: Efficiently training massive language models

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T11:21:19.753237Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-18T11:17:08.108565Z digest=sha256:0e6ae8f457c72426eebc660db4ad388e4cfedce4db23c74ed803f410951aaa5a

Observation 0c1a4d65-a678-4101-b51a-904b5a89482b · outbound

This paper cites How chatgpt works: A look inside large language models.

TrustLLM: Trustworthiness in Large Language Models How chatgpt works: A look inside large language models

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T11:21:19.771494Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-18T11:17:08.108565Z digest=sha256:9e31f43a9fc14fcfa62dc91b1a2f09751b5f0dce7b749c4c4c4a901d9ad8c478

Observation 98d53138-35f9-41de-8164-7891decd2683 · outbound

This paper cites LoRA: Low-Rank Adaptation of Large Language Models.

TrustLLM: Trustworthiness in Large Language Models LoRA: Low-Rank Adaptation of Large Language Models

Reference 39

Resolution
verified exact
local_arxiv, observed 2026-05-18T11:17:08.320405Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-18T11:17:08.108565Z digest=sha256:c214520ae543c8b8acf4a6f4d5e93b51d2d291883ed9bbf2fa7b803a29ec6ef6

Observation d56339b2-3d20-40a6-b85a-798af7e6a869 · outbound

This paper cites QLoRA: Efficient Finetuning of Quantized LLMs.

TrustLLM: Trustworthiness in Large Language Models QLoRA: Efficient Finetuning of Quantized LLMs

Reference 40

Resolution
verified exact
local_arxiv, observed 2026-05-18T11:17:08.326899Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-18T11:17:08.108565Z digest=sha256:babfbf21b5e992907eb8fd5d516ecefb5153a3f26c254654c73c0d67c5e2976c

Observation a02c6d68-7df4-4868-a16a-49aec21dd956 · outbound

This paper cites Pathways: Asynchronous distributed dataflow for ml.

TrustLLM: Trustworthiness in Large Language Models Pathways: Asynchronous distributed dataflow for ml

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T11:21:20.033175Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-18T11:17:08.108565Z digest=sha256:c1b7716b3cc9cde76823b73d70c76694fe570b363923a8b9d7698ffa93cfefe3

Observation 54bbc698-5eed-491c-933b-b877aa9833ed · outbound

This paper cites Ai alignment: A comprehensive survey.

TrustLLM: Trustworthiness in Large Language Models Ai alignment: A comprehensive survey

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T11:21:19.814967Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-18T11:17:08.108565Z digest=sha256:005386b52c03a34d0d801edf424f1716538e3a138077b756c871da32d77baee5

Observation 4dc02eac-8e4e-4c7e-baaa-179c0c1033a5 · outbound

This paper cites Training language models to follow instructions with human feedback.

TrustLLM: Trustworthiness in Large Language Models Training language models to follow instructions with human feedback

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T11:21:19.658553Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-18T11:17:08.108565Z digest=sha256:aa55d200fdb7a40451337bc5e4e23026f92c889d74b3df37f6c90c992fc33082

Observation 07607afd-2f0d-40dd-8792-782b8dcb8444 · outbound

This paper cites Improving Language Model Negotiation with Self-Play and In-Context Learning from AI Feedback.

TrustLLM: Trustworthiness in Large Language Models Improving Language Model Negotiation with Self-Play and In-Context Learning from AI Feedback

Reference 44

Resolution
verified exact
arxiv_id, observed 2026-05-18T11:17:08.334483Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-18T11:17:08.108565Z digest=sha256:d1097517ce40ee0dfe87e01c6d34083caf370f514fc29d3ccd9f353c052bdf81

Observation c9b0aad9-9dd6-45c0-ade3-e73ac5a1394c · outbound

This paper cites Principle-Driven Self-Alignment of Language Models from Scratch with Minimal Human Supervision.

TrustLLM: Trustworthiness in Large Language Models Principle-Driven Self-Alignment of Language Models from Scratch with Minimal Human Supervision

Reference 45

Resolution
verified exact
arxiv_id, observed 2026-05-18T11:17:08.340723Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-18T11:17:08.108565Z digest=sha256:5ec925698cc1344c20923f9d3e7ebd4ac47e54d18a2ca13015e2f39431c3247e

Observation 093f81f9-5948-41af-a9ea-7bd66ffe83fb · outbound

This paper cites Rl4f: Generating natural language feedback with reinforcement learning for repairing model outputs.

TrustLLM: Trustworthiness in Large Language Models Rl4f: Generating natural language feedback with reinforcement learning for repairing model outputs

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T11:21:19.840964Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-18T11:17:08.108565Z digest=sha256:cf4ea65b247f6b282cd7b4c255cc84043aa4d12ce9bb3ef56468562c91d3772d

Observation 8601443d-1632-4d69-907c-c0eeff835971 · outbound

This paper cites Measuring Progress on Scalable Oversight for Large Language Models.

TrustLLM: Trustworthiness in Large Language Models Measuring Progress on Scalable Oversight for Large Language Models

Reference 47

Resolution
verified exact
local_arxiv, observed 2026-05-18T11:17:08.347452Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-18T11:17:08.108565Z digest=sha256:cca2a141912accfd45d46f2eced8a962e512bb38619d571c3c8881c904875418

Observation 2350b892-3ea6-4f1e-9b5f-bc481c8eb5c7 · outbound

This paper cites Discovering Language Model Behaviors with Model-Written Evaluations.

TrustLLM: Trustworthiness in Large Language Models Discovering Language Model Behaviors with Model-Written Evaluations

Reference 48

Resolution
verified exact
local_arxiv, observed 2026-05-18T11:17:08.353781Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-18T11:17:08.108565Z digest=sha256:19fa36eb61ccf0cbc3f967f9b9554e38b504b98c335c6356d8acec5b69a6139b

Observation 3b971218-5d04-47b2-94ab-5ff3d50aec18 · outbound

This paper cites Improving Factuality and Reasoning in Language Models through Multiagent Debate.

TrustLLM: Trustworthiness in Large Language Models Improving Factuality and Reasoning in Language Models through Multiagent Debate

Reference 49

Resolution
verified exact
local_arxiv, observed 2026-05-18T11:17:08.361185Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-18T11:17:08.108565Z digest=sha256:fbcea066aedcdbb1c0412db956f8f330f48d10739f322a6dcd5be877be01bdea

Observation 1822c6ec-5d9d-41b8-aa9e-7213115c1b13 · outbound

This paper cites Characterizing Manipulation from AI Systems.

TrustLLM: Trustworthiness in Large Language Models Characterizing Manipulation from AI Systems

Reference 50

Resolution
verified exact
arxiv_id, observed 2026-05-18T11:17:08.368700Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-18T11:17:08.108565Z digest=sha256:aa67a456611cab3bb94ab2b9ea9c6f93a6f97b8398c58a53739512858db7d05a

Observation 841f2e03-f5ff-4121-9b72-9b295046ea61 · outbound

This paper cites RLAIF vs. RLHF: Scaling Reinforcement Learning from Human Feedback with AI Feedback.

TrustLLM: Trustworthiness in Large Language Models RLAIF vs. RLHF: Scaling Reinforcement Learning from Human Feedback with AI Feedback

Reference 51

Resolution
verified exact
local_arxiv, observed 2026-05-18T11:17:08.374817Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-18T11:17:08.108565Z digest=sha256:63aad77def888274c9bdbd104b9ae4f1f856a358a367298a3155412aedb4a537

Observation a8cf28a0-091a-425a-9a68-f21878a9f399 · outbound

This paper cites A Generalist Agent.

TrustLLM: Trustworthiness in Large Language Models A Generalist Agent

Reference 52

Resolution
verified exact
local_arxiv, observed 2026-05-18T11:17:08.379865Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-18T11:17:08.108565Z digest=sha256:421bb7821358cc4fcece65d430e612574673c77397f769c571d2c6d5cc31bcf8

Observation fa881004-3ec9-458d-8ff8-1326f8aa3783 · outbound

This paper cites Constitutional AI: Harmlessness from AI Feedback.

TrustLLM: Trustworthiness in Large Language Models Constitutional AI: Harmlessness from AI Feedback

Reference 53

Resolution
verified exact
local_arxiv, observed 2026-05-18T11:17:08.385786Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-18T11:17:08.108565Z digest=sha256:fe6d78bdf9c2242170375a3ff677c7c6fbd676c4ba4724037dcc3856b0717249

Observation d2077b70-7dd0-4061-be6d-d022331164c4 · outbound

This paper cites The Effects of Reward Misspecification: Mapping and Mitigating Misaligned Models.

TrustLLM: Trustworthiness in Large Language Models The Effects of Reward Misspecification: Mapping and Mitigating Misaligned Models

Reference 54

Resolution
verified exact
arxiv_id, observed 2026-05-21T07:08:52.819373Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-18T11:17:08.108565Z digest=sha256:a028946dcf4f626d0b9c5f0b016f9c7999b84ae11828e7fb1307649c44c27d58

Observation 3bea1e0b-b33b-462a-88fd-627d9f23d3b0 · outbound

This paper cites Cooperative inverse reinforcement learning.

TrustLLM: Trustworthiness in Large Language Models Cooperative inverse reinforcement learning

Reference 55

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T11:21:19.642597Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-18T11:17:08.108565Z digest=sha256:c78a4f9064715777ab030bbdff3390036ea13f457ac96b547d2a5e76480c3599

Observation 2f1e667a-5d98-4dcb-803c-b1a9ab60d61a · outbound

This paper cites Survey of hallucination in natural language generation.

TrustLLM: Trustworthiness in Large Language Models Survey of hallucination in natural language generation

Reference 56

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T11:21:19.855907Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-18T11:17:08.108565Z digest=sha256:4840c4b4aa643c54a65a64b576f65a8e00ed25c37f67679e2460617f29b25bfd

Observation 03a4e4f7-eee3-4ce1-ab81-fa6886a4a138 · outbound

This paper cites A Survey on Hallucination in Large Language Models: Principles, Taxonomy, Challenges, and Open Questions.

TrustLLM: Trustworthiness in Large Language Models A Survey on Hallucination in Large Language Models: Principles, Taxonomy, Challenges, and Open Questions

Reference 57

Resolution
verified exact
local_arxiv, observed 2026-05-18T11:17:08.397348Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-18T11:17:08.108565Z digest=sha256:3e77ac6aa137ec1dffb1e2ea304a1d53a0cd50b3e4b2fa5aaafecfa3a2138cae

Observation 09964268-a039-461c-a507-76ffabe403b8 · outbound

This paper cites Factuality challenges in the era of large language models.

TrustLLM: Trustworthiness in Large Language Models Factuality challenges in the era of large language models

Reference 58

Resolution
verified exact
arxiv_id, observed 2026-05-18T11:17:08.404580Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-18T11:17:08.108565Z digest=sha256:b64e6c867025c35765d1f8b45b701045d21a4c1b4ceb7b108f0ef0272ec8834d

Observation f0fe0f8b-5f0a-49bf-88e5-4a01939f063c · outbound

This paper cites Combating Misinformation in the Age of LLMs: Opportunities and Challenges.

TrustLLM: Trustworthiness in Large Language Models Combating Misinformation in the Age of LLMs: Opportunities and Challenges

Reference 59

Resolution
verified exact
arxiv_id, observed 2026-05-18T11:17:08.409989Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-18T11:17:08.108565Z digest=sha256:2278157a487e54fc72166b5a14f6c283078b36b0e3d34081fdc392f08172a79c

Observation c8c59fef-2396-4038-96ac-b3eb3c3154a5 · outbound

This paper cites 10 ways cybercriminals can abuse large language models.

TrustLLM: Trustworthiness in Large Language Models 10 ways cybercriminals can abuse large language models

Reference 60

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T11:21:19.870647Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-18T11:17:08.108565Z digest=sha256:bd02cd896c63b7aa7b0c893da055aa71c52a89283d7b0a9f2ee5c2ffbca2bf0b

Observation 8489c251-0396-4469-8ca0-2758d68716f8 · outbound

This paper cites Jailbroken: How Does LLM Safety Training Fail?.

TrustLLM: Trustworthiness in Large Language Models Jailbroken: How Does LLM Safety Training Fail?

Reference 61

Resolution
verified exact
local_arxiv, observed 2026-05-18T11:17:08.415633Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-18T11:17:08.108565Z digest=sha256:73a4907b52da30cea2a7227b49759aeffa811f1df1e36ede3592dc8c48758782

Observation 4a2094e0-2a25-4150-8d07-36e430454598 · outbound

This paper cites Unraveling the link between translations and gender bias in llms.

TrustLLM: Trustworthiness in Large Language Models Unraveling the link between translations and gender bias in llms

Reference 62

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T11:21:19.764818Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-18T11:17:08.108565Z digest=sha256:65b2fa4e577649ad8f0270acb3d333e898343752616ca1b21d8bf86dcb450f49

Observation 24436732-944f-4a9e-9718-1066e8b8c3a5 · outbound

This paper cites Navigating the biases in llm generative ai: A guide to responsible implementation.

TrustLLM: Trustworthiness in Large Language Models Navigating the biases in llm generative ai: A guide to responsible implementation

Reference 63

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T11:21:19.778614Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-18T11:17:08.108565Z digest=sha256:b8d04b8230fcc09ea96c761570b7f87f011f9b842a3d457b090d559419177853

Observation 7562b62f-f0be-4df5-99d8-006de124854d · outbound

This paper cites Large language models may leak personal data.

TrustLLM: Trustworthiness in Large Language Models Large language models may leak personal data

Reference 64

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T11:21:19.652044Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-18T11:17:08.108565Z digest=sha256:c12c0c5925f8025524a9bee49b3660c4659abccd1495ea98dcae7718ac92d0ff

Observation 225241b2-9bf9-4567-89e7-280534c4f6ef · outbound

This paper cites Deid-gpt: Zero-shot medical text de-identification by gpt-4.

TrustLLM: Trustworthiness in Large Language Models Deid-gpt: Zero-shot medical text de-identification by gpt-4

Reference 65

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T11:21:19.881445Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-18T11:17:08.108565Z digest=sha256:4399adc6c952ddb40c778a83c884a2a53864dcecd63b0c2b7249b2543c0a21c0

Observation 4a987fff-772a-4891-9db9-28be8bf80211 · outbound

This paper cites What does it mean to align ai with human values?.

TrustLLM: Trustworthiness in Large Language Models What does it mean to align ai with human values?

Reference 66

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T11:21:19.639401Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-18T11:17:08.108565Z digest=sha256:8a0ac930b7cd0d20b4e8f86977dfa2f2e49f59d2a9d188b207f07c82b5fb3b38

Observation 54305edd-5d14-460b-8ee5-a232b9131252 · outbound

This paper cites Openai.

TrustLLM: Trustworthiness in Large Language Models Openai

Reference 67

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T11:21:19.721833Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-18T11:17:08.108565Z digest=sha256:084250b32caed61350a148ac5909cda8f5f0ec6b02c69e8c967a3a9745bd9cbd

Observation 1d3f8eda-4cf1-49cc-87aa-dfcba66f9987 · outbound

This paper cites Ai at meta.

TrustLLM: Trustworthiness in Large Language Models Ai at meta

Reference 68

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T11:21:19.795110Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-18T11:17:08.108565Z digest=sha256:d897323efcb3529119832d9977f84aec91d58883a0a8910747fc22eea706264c

Observation f2604d1a-0cf0-41bc-929c-a2f50db9d375 · outbound

This paper cites Llama 2: Open Foundation and Fine-Tuned Chat Models.

TrustLLM: Trustworthiness in Large Language Models Llama 2: Open Foundation and Fine-Tuned Chat Models

Reference 69

Resolution
verified exact
local_arxiv, observed 2026-05-18T11:17:08.422349Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-18T11:17:08.108565Z digest=sha256:0ac7c1504f64094678acadc7f3ec43e208f6e1913dafa0cdeca19e02b8aae78a

Observation dacd4082-fada-42c9-bcd6-eccb45ff3db4 · outbound

This paper cites Holistic Evaluation of Language Models.

TrustLLM: Trustworthiness in Large Language Models Holistic Evaluation of Language Models

Reference 70

Resolution
verified exact
local_arxiv, observed 2026-05-18T11:17:08.428332Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-18T11:17:08.108565Z digest=sha256:c55711a2eaa2b71397c233dbd4c00355ea21bb43914dde016ad45fb27e523b0b

Observation 7133d9f0-4fd0-4606-bbf8-076a4c12371c · outbound

This paper cites DecodingTrust: A Comprehensive Assessment of Trustworthiness in GPT Models.

TrustLLM: Trustworthiness in Large Language Models DecodingTrust: A Comprehensive Assessment of Trustworthiness in GPT Models

Reference 71

Resolution
verified exact
arxiv_id, observed 2026-05-18T11:17:08.434626Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-18T11:17:08.108565Z digest=sha256:aae986eae5ddb8f67b6c0f30acdd705fad83cae347b1a2cc1e22b037f23b1326

Observation 5c4d3f45-cc05-44bc-8fc8-e190eb8f5f1b · outbound

This paper cites Trustworthy LLMs: a Survey and Guideline for Evaluating Large Language Models' Alignment.

TrustLLM: Trustworthiness in Large Language Models Trustworthy LLMs: a Survey and Guideline for Evaluating Large Language Models' Alignment

Reference 72

Resolution
verified exact
local_arxiv, observed 2026-05-18T11:17:08.440377Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-18T11:17:08.108565Z digest=sha256:69e8c80dabbad80885c0824c5463a1ecc93733bc4dcd49bb54011e9e8b0acc2a

Observation f80178cc-aec5-402b-b6c5-a92bbc41ba05 · outbound

This paper cites Do-Not-Answer: A Dataset for Evaluating Safeguards in LLMs.

TrustLLM: Trustworthiness in Large Language Models Do-Not-Answer: A Dataset for Evaluating Safeguards in LLMs

Reference 73

Resolution
metadata mismatch
arxiv_id, observed 2026-05-18T11:17:08.446480Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-18T11:17:08.108565Z digest=sha256:8a5a218e0bb611623310cdbe9c5cc2d74cd392b4c98465d238abcc32bb1df8d8

Observation 9dd37d76-2f6c-4b01-8e03-63adfa5b6950 · outbound

This paper cites Chatbot arena leaderboard week 8: Introducing mt-bench and vicuna-33b.

TrustLLM: Trustworthiness in Large Language Models Chatbot arena leaderboard week 8: Introducing mt-bench and vicuna-33b

Reference 74

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T11:21:19.673179Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-18T11:17:08.108565Z digest=sha256:9a115f9f17d3d91f428d64090d64c98c344a0afa8c59c1c3727f811a4c5608a6

Observation 13bdfdcf-5d0a-4771-b1e6-a3b37657acb4 · outbound

This paper cites The big benchmarks collection - a open-llm-leaderboard collection.

TrustLLM: Trustworthiness in Large Language Models The big benchmarks collection - a open-llm-leaderboard collection

Reference 75

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T11:21:19.696116Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-18T11:17:08.108565Z digest=sha256:08dd4bead71753c94d590ddadced38deb92aff972ee2b3aebe3c95af26190051

Observation fcbfa731-8aa7-40e6-bc1b-4b1e1dc45dcd · outbound

This paper cites https://platform.openai.com/docs/guides/moderation.

TrustLLM: Trustworthiness in Large Language Models https://platform.openai.com/docs/guides/moderation

Reference 76

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T11:21:19.655349Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-18T11:17:08.108565Z digest=sha256:941edf8df637cbd12b014e0b87e44ecf0e617ebdbc270a3971a52a5057c35e2c

Observation 9ed9d9d5-4405-4513-8741-6ae38f48e506 · outbound

This paper cites The foundation model transparency index.

TrustLLM: Trustworthiness in Large Language Models The foundation model transparency index

Reference 77

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T11:21:19.894565Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-18T11:17:08.108565Z digest=sha256:54cc5167aee7c852fd6ccb6c8761c70c5b8435fbb964000fb4f18daace5d4ad7

Observation c4509149-f6c5-49f1-936c-cfd39c17430b · outbound

This paper cites an unresolved cited work.

TrustLLM: Trustworthiness in Large Language Models Unresolved cited work

Reference 78

Resolution
unresolved
raw_fallback, observed 2026-05-18T11:21:19.707646Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-18T11:17:08.108565Z digest=sha256:63af4a3d88aefaf0fb3b43fe05cc2620fa00b8c8870cbe9aba3e7bf3b78be07d

Observation 3597cca9-b5aa-4aeb-af41-4871c3096910 · outbound

This paper cites Ernie - baidu yiyan.

TrustLLM: Trustworthiness in Large Language Models Ernie - baidu yiyan

Reference 79

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T11:21:19.710505Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-18T11:17:08.108565Z digest=sha256:d4024093c3227621858d310a24a179d4c9337b90cfbcbc5181c1fed5cd24d8c1

Observation e90a6ad9-c56f-4b54-bf1a-dcc3730ea6ce · outbound

This paper cites Attention is all you need.

TrustLLM: Trustworthiness in Large Language Models Attention is all you need

Reference 80

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T11:21:19.636110Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-18T11:17:08.108565Z digest=sha256:2ac5f242afe707e37f16e482f56b5a8b5499b2a4bd298637db7b3bf936716fe7

Observation 41418d48-616d-453c-ab8e-8c92adccc940 · outbound

This paper cites an open assistant for everyone by laion.

TrustLLM: Trustworthiness in Large Language Models an open assistant for everyone by laion

Reference 81

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T11:21:19.781764Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-18T11:17:08.108565Z digest=sha256:fdab1869a6f0ce402c8ee325d314dcc54d4098d31cad3523e05647b795c79234

Observation 76ebbf8d-a854-40ba-a1dd-f20d9205b14b · outbound

This paper cites Gonzalez, Ion Stoica, and Eric P.

TrustLLM: Trustworthiness in Large Language Models Gonzalez, Ion Stoica, and Eric P

Reference 82

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T11:21:19.744685Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-18T11:17:08.108565Z digest=sha256:c59e6a4f74a9151c1fb1c8af9bac7c529a212758a16b823fb671fde9a1bf9a91

Observation 6c35a320-bdd0-45f5-a6da-950e1bdc8298 · outbound

This paper cites https://nvlpubs.nist.gov/nistpubs/ ai/NIST.AI.100-1.pdf.

TrustLLM: Trustworthiness in Large Language Models https://nvlpubs.nist.gov/nistpubs/ ai/NIST.AI.100-1.pdf

Reference 83

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T11:21:19.788422Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-18T11:17:08.108565Z digest=sha256:1731bf9cff44ebbf87116b44123e3157fc571aa5f55fec3918f2a7bbee6cafa8

Observation 99e38ce1-b67b-4fb4-83ac-ed546aa32f7c · outbound

This paper cites Enron email dataset.

TrustLLM: Trustworthiness in Large Language Models Enron email dataset

Reference 84

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T11:21:19.645256Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-18T11:17:08.108565Z digest=sha256:0850020a53ba73403cc50eb35f6606f55bc41a29e52f21122c5c4896f3ee9685

Observation acc153dc-42c9-48d5-96b3-e0226b5d6543 · outbound

This paper cites Guest editors’ introduction: machine ethics.

TrustLLM: Trustworthiness in Large Language Models Guest editors’ introduction: machine ethics

Reference 85

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T11:21:19.749058Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-18T11:17:08.108565Z digest=sha256:5ae9197c98b8bc964721bb65c3c59984574c67ed4d9a1d54abeaa7d64338b3da

Observation be040ea0-e057-4426-9cf3-06fa61806efe · outbound

This paper cites Machine ethics: Creating an ethical intelligent agent.

TrustLLM: Trustworthiness in Large Language Models Machine ethics: Creating an ethical intelligent agent

Reference 86

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T11:21:19.884697Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-18T11:17:08.108565Z digest=sha256:ef18fca7912bff563762e7ca723204ed1120afa36cecb9da08f1eb1812bdfa35

Observation 02025677-03fd-4f7c-859c-9c8b52f261cf · outbound

This paper cites Emergent Abilities of Large Language Models.

TrustLLM: Trustworthiness in Large Language Models Emergent Abilities of Large Language Models

Reference 87

Resolution
verified exact
local_arxiv, observed 2026-05-18T11:17:08.451351Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-18T11:17:08.108565Z digest=sha256:ad8840dd5126261388dfb06527209609a7fd06c359704d0bce9976d28b0aa9b9

Observation 7ae3047f-8d4d-42e6-bb11-203122fbb27c · outbound

This paper cites Chain-of-thought prompting elicits reasoning in large language models.

TrustLLM: Trustworthiness in Large Language Models Chain-of-thought prompting elicits reasoning in large language models

Reference 88

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T11:21:20.025829Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-18T11:17:08.108565Z digest=sha256:2afcb3857eb7223bdefa291381db4f7d5c31cf3fe510d757925e36ff842835fa

Observation bba74d8d-1ca0-4110-aac3-e8bd91586c95 · outbound

This paper cites Scaling Instruction-Finetuned Language Models.

TrustLLM: Trustworthiness in Large Language Models Scaling Instruction-Finetuned Language Models

Reference 89

Resolution
verified exact
local_arxiv, observed 2026-05-18T11:17:08.456495Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-18T11:17:08.108565Z digest=sha256:733b574ea71cfdf1e959ce588d78f3b108711db028e86ab6fb1a03c373a6c04f

Observation bcb2691c-cb97-4b6d-8b9e-21ac98881482 · outbound

This paper cites Exploring the limits of transfer learning with a unified text-to-text transformer.

TrustLLM: Trustworthiness in Large Language Models Exploring the limits of transfer learning with a unified text-to-text transformer

Reference 90

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T11:21:19.703920Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-18T11:17:08.108565Z digest=sha256:c99b8189573b310702a18efbab3915269f77455823aa9fc41ffaa25fdd083ee3

Observation fa1c89e3-76a9-47a9-9493-e4290ccf391b · outbound

This paper cites Scaling Laws for Neural Language Models.

TrustLLM: Trustworthiness in Large Language Models Scaling Laws for Neural Language Models

Reference 91

Resolution
verified exact
local_arxiv, observed 2026-05-18T11:17:08.463255Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-18T11:17:08.108565Z digest=sha256:1586b6a820ae2dbe3040a62f0396a19a25d0d47cd4f7d4c755589faec21558ab

Observation 576152b0-6f81-4d4b-97a0-89d797b42952 · outbound

This paper cites Training Compute-Optimal Large Language Models.

TrustLLM: Trustworthiness in Large Language Models Training Compute-Optimal Large Language Models

Reference 92

Resolution
verified exact
local_arxiv, observed 2026-05-18T11:17:08.467578Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-18T11:17:08.108565Z digest=sha256:da6c6253c24cccded6355c2e3b5861f53bc1c47650b0645e2a22ce8cb78d0043

Observation 441eeb77-e4c6-4f1d-8066-b312a393e33b · outbound

This paper cites Proximal Policy Optimization Algorithms.

TrustLLM: Trustworthiness in Large Language Models Proximal Policy Optimization Algorithms

Reference 93

Resolution
verified exact
local_arxiv, observed 2026-05-18T11:17:08.477326Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-18T11:17:08.108565Z digest=sha256:e135958a0ff1bf635716e313c637a94bc1d13b9a755a8b56d2245a58e3210314

Observation 42c92d63-24f4-4771-ad82-0777d5589509 · outbound

This paper cites RAFT: Reward rAnked FineTuning for Generative Foundation Model Alignment.

TrustLLM: Trustworthiness in Large Language Models RAFT: Reward rAnked FineTuning for Generative Foundation Model Alignment

Reference 94

Resolution
verified exact
local_arxiv, observed 2026-05-18T11:17:08.482758Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-18T11:17:08.108565Z digest=sha256:7278a28ba4b276704ccbeee73c70b4ad1b3eadfc174b5e7f62f7555596f14036

Observation 45b8b23c-2054-42df-bf3e-bfb40f708d26 · outbound

This paper cites OpenChat: Advancing Open-source Language Models with Mixed-Quality Data.

TrustLLM: Trustworthiness in Large Language Models OpenChat: Advancing Open-source Language Models with Mixed-Quality Data

Reference 95

Resolution
verified exact
arxiv_id, observed 2026-05-18T11:17:08.488856Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-18T11:17:08.108565Z digest=sha256:091ba7f676baa17a39e873f5f937576a3ab6f8c9768009078379dfb3cde8b521

Observation 51f31bb9-f395-427e-8558-ebef725cee31 · outbound

This paper cites Chain of Hindsight Aligns Language Models with Feedback.

TrustLLM: Trustworthiness in Large Language Models Chain of Hindsight Aligns Language Models with Feedback

Reference 96

Resolution
verified exact
arxiv_id, observed 2026-05-18T11:17:08.494101Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-18T11:17:08.108565Z digest=sha256:c09ebd191b61016096b3a91e7f2076bd853be5246b93b67b7efc7ea21b895abf

Observation 037599fd-e5a9-47d5-9b37-633391134655 · outbound

This paper cites Training Socially Aligned Language Models on Simulated Social Interactions.

TrustLLM: Trustworthiness in Large Language Models Training Socially Aligned Language Models on Simulated Social Interactions

Reference 97

Resolution
metadata mismatch
arxiv_id, observed 2026-05-18T11:17:08.499802Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-18T11:17:08.108565Z digest=sha256:f8400953b5f6367a24a69a6ac7154e7df966f7aa322786ee2fcf9632751960fc

Observation c1f558d4-6c9c-445e-b5c6-483ad92e0747 · outbound

This paper cites Yu, Qiang Yang, and Xing Xie.

TrustLLM: Trustworthiness in Large Language Models Yu, Qiang Yang, and Xing Xie

Reference 98

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T11:21:19.811752Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-18T11:17:08.108565Z digest=sha256:4f1a2f5a2ee0cac4ef8c4d0d68e466a3419494788a5f9664c4bc773882e4aee0

Observation a368994f-6309-4724-903f-16d32c1729ee · outbound

This paper cites Can chatgpt forecast stock price movements? return pre- dictability and large language models.

TrustLLM: Trustworthiness in Large Language Models Can chatgpt forecast stock price movements? return pre- dictability and large language models

Reference 99

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T11:21:19.665723Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-18T11:17:08.108565Z digest=sha256:4c63f8707ef8d91c621c8cfaecf3dc51601dac955ec6a06c000230e5e02f7d44

Observation 370d8f83-950f-40dd-9c0f-e5224a579b7e · outbound

This paper cites Sentiment analysis in the era of large language models: A reality check.

TrustLLM: Trustworthiness in Large Language Models Sentiment analysis in the era of large language models: A reality check

Reference 100

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T11:17:08.839813Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-18T11:17:08.108565Z digest=sha256:b02900d40150c8e6d96e0c3ee09632752cf0d7da2c90741eafed426d9a2fb6fb

Pith citing papers

Observation e5944dee-3fd8-4fc3-9260-08aec534976f · inbound

Large Language Models: A Survey cites this paper.

Large Language Models: A Survey TrustLLM: Trustworthiness in Large Language Models

Reference 221

Resolution
verified exact
arxiv_id, observed 2026-05-18T11:17:09.209714Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-11T15:22:54.023279Z digest=sha256:6fc8d7f1f9cd3cded6a4636dc4b2cb35c3fabcf559294a5653861f95bfebf730

Observation 8853d8da-4b81-4463-9835-b3e09969d32c · inbound

JailbreakBench: An Open Robustness Benchmark for Jailbreaking Large Language Models cites this paper.

JailbreakBench: An Open Robustness Benchmark for Jailbreaking Large Language Models TrustLLM: Trustworthiness in Large Language Models

Reference 49

Resolution
verified exact
arxiv_id, observed 2026-05-18T11:17:09.209714Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-15T06:08:05.386345Z digest=sha256:ef9edcd7b423dbc1e64a166898ee51c6d45975aecb6bcb7e2dbda35f8dc6e38f

Observation 7de9155a-012d-4aa1-8dcb-1ebcc44a60de · inbound

Trustworthiness in Retrieval-Augmented Generation Systems: A Survey cites this paper.

Trustworthiness in Retrieval-Augmented Generation Systems: A Survey TrustLLM: Trustworthiness in Large Language Models

Reference 20

Resolution
verified exact
local_arxiv, observed 2026-05-23T21:08:25.836081Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-23T21:08:11.787013Z digest=sha256:eb69d9f4d7293c58bec8ceab354c2adef4c9cd8d0d35ba1561e8a34ad181d378

Observation 46c76b80-b86d-4d6d-b0ab-a93d47e301be · inbound

Entry-level guide to the use of large language models for medical research cites this paper.

Entry-level guide to the use of large language models for medical research TrustLLM: Trustworthiness in Large Language Models

Reference 13

Resolution
verified exact
local_arxiv, observed 2026-05-23T19:28:21.486508Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-23T19:26:48.548573Z digest=sha256:5a7ab1e4305e02e78d0ac2d3978052f1dee0dc3b52c8d7004a54648f9202cc3d

Observation bd754c48-a272-4f69-83ef-dd0af4ff3f1b · inbound

Opportunities and Challenges of Large Language Models for Low-Resource Languages in Humanities Research cites this paper.

Opportunities and Challenges of Large Language Models for Low-Resource Languages in Humanities Research TrustLLM: Trustworthiness in Large Language Models

Reference 58

Resolution
metadata mismatch
local_arxiv, observed 2026-05-23T16:38:11.367395Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-23T16:36:02.895613Z digest=sha256:b8c73b20556ebd40a2f1ed1bfe89a39b7038404c76bc3af41e1f1d35ec77cbc0

Observation 2dd88203-ff68-4191-a56e-9eea7459501c · inbound

Enabling Global, Human-Centered Explanations for LLMs:From Tokens to Interpretable Code and Test Generation cites this paper.

Enabling Global, Human-Centered Explanations for LLMs:From Tokens to Interpretable Code and Test Generation TrustLLM: Trustworthiness in Large Language Models

Reference 60

Resolution
verified exact
local_arxiv, observed 2026-05-22T23:42:16.044178Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-22T23:41:22.017848Z digest=sha256:bdd6108974fdbae5f9b1ae2b16432f7a1080833e961f5705c3bfbd5ac0ef82eb

Observation f306397b-db9e-45bd-bd57-23b43b64ac12 · inbound

ReGA: Model-Based Safeguard for LLMs via Representation-Guided Abstraction cites this paper.

ReGA: Model-Based Safeguard for LLMs via Representation-Guided Abstraction TrustLLM: Trustworthiness in Large Language Models

Reference 15

Resolution
verified exact
local_arxiv, observed 2026-05-19T11:37:15.973880Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-19T11:34:09.428653Z digest=sha256:8a45e157518f3688a29e5098c250a67ed882deac0293fcae8aa4ab7e9a2dffa2

Observation fdea7240-6eda-4d2a-9986-f2f0fe0dec68 · inbound

Downgrade to Upgrade: Optimizer Simplification Enhances Robustness in LLM Unlearning cites this paper.

Downgrade to Upgrade: Optimizer Simplification Enhances Robustness in LLM Unlearning TrustLLM: Trustworthiness in Large Language Models

Reference 6

Resolution
verified exact
arxiv_id, observed 2026-05-18T11:17:09.209714Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-18T10:44:53.516653Z digest=sha256:df85445d9fde4250eff79e4e00b75b070ab3c3d58fe30054c153257fd565e4a5

Observation d7032d39-2799-4dd7-aa96-b2a198868bcf · inbound

Erase to Improve: Erasable Reinforcement Learning for Search-Augmented LLMs cites this paper.

Erase to Improve: Erasable Reinforcement Learning for Search-Augmented LLMs TrustLLM: Trustworthiness in Large Language Models

Reference 5

Resolution
verified exact
arxiv_id, observed 2026-05-18T11:17:09.209714Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-18T11:06:20.058342Z digest=sha256:8cdee9c3b65564e98df22c8c10169f7f0dfa5157b9d9c8ecf94d1777c817c65f

Observation bbf26d3a-a66b-4ed6-87db-0fc1853522d2 · inbound

Leak@$k$: Unlearning Does Not Make LLMs Forget Under Probabilistic Decoding cites this paper.

Leak@$k$: Unlearning Does Not Make LLMs Forget Under Probabilistic Decoding TrustLLM: Trustworthiness in Large Language Models

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-03T23:39:09.106641Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T23:39:09.106641Z digest=sha256:f84b726b355a42d04b1099c284b0272cbf9ad73098b393e61a69692cebc39891

Observation e752d3d0-1aee-489e-b3ed-8273c6d81d6f · inbound

OutSafe-Bench: A Benchmark for Multimodal Offensive Content Detection in Large Language Models cites this paper.

OutSafe-Bench: A Benchmark for Multimodal Offensive Content Detection in Large Language Models TrustLLM: Trustworthiness in Large Language Models

Reference 51

Resolution
verified exact
arxiv_id, observed 2026-05-18T11:17:09.209714Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-17T22:29:36.960961Z digest=sha256:8568eb33487297f14e5491366dd2135d6bb54e5fa76591aa6a3dc39e58ab60f7

Observation ad293aa8-23ec-4169-aedc-2ca13de390d0 · inbound

On the Factual Consistency of Text-based Explainable Recommendation Models cites this paper.

On the Factual Consistency of Text-based Explainable Recommendation Models TrustLLM: Trustworthiness in Large Language Models

Reference 10

Resolution
metadata mismatch
local_arxiv, observed 2026-05-21T15:50:19.453503Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-21T15:45:20.394551Z digest=sha256:a475c93e0218c085b8e8af53db92fb1313127485bd262bf984a1bb44cd7fc74c

Observation 4bb73bbe-4165-45a5-aa9b-f8ec8177015d · inbound

SafeCRS: Personalized Safety Alignment for LLM-Based Conversational Recommender Systems cites this paper.

SafeCRS: Personalized Safety Alignment for LLM-Based Conversational Recommender Systems TrustLLM: Trustworthiness in Large Language Models

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-02T19:10:40.562289Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T19:10:40.562289Z digest=sha256:882031f657d7d1068366ed8b5e1cd86bf7482a2c22acd5de9781413315df7e4d

Observation 5b64b74b-f1f4-46a0-a682-64ca777b4076 · inbound

Shorter, but Still Trustworthy? An Empirical Study of Chain-of-Thought Compression cites this paper.

Shorter, but Still Trustworthy? An Empirical Study of Chain-of-Thought Compression TrustLLM: Trustworthiness in Large Language Models

Reference 9

Resolution
verified exact
arxiv_id, observed 2026-05-18T11:17:09.209714Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-13T17:16:47.284564Z digest=sha256:0e3d8bb861d738178c2ce229f4bbc1b77cef21fadc5bf7abc9ae74e621f325ca

Observation 0e3ea996-6cc6-4b19-bfba-bb2e154434d3 · inbound

Guardian-as-an-Advisor: Advancing Next-Generation Guardian Models for Trustworthy LLMs cites this paper.

Guardian-as-an-Advisor: Advancing Next-Generation Guardian Models for Trustworthy LLMs TrustLLM: Trustworthiness in Large Language Models

Reference 26

Resolution
verified exact
arxiv_id, observed 2026-05-18T11:17:09.209714Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-10T17:27:13.339411Z digest=sha256:65042a199dec3d86716bcce008ede03def50b3eefcccb789afef739ef74d36ab

Observation d5b68b55-d08d-430d-9674-b90a026ffb27 · inbound

AVID: A Benchmark for Omni-Modal Audio-Visual Inconsistency Understanding via Agent-Driven Construction cites this paper.

AVID: A Benchmark for Omni-Modal Audio-Visual Inconsistency Understanding via Agent-Driven Construction TrustLLM: Trustworthiness in Large Language Models

Reference 16

Resolution
verified exact
arxiv_id, observed 2026-05-18T11:17:09.209714Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-10T12:23:56.510443Z digest=sha256:d525f8881257ce92afb544366b755f585f0bd2cc2aed6e81968d629205a910a2

Observation 21b38b88-2462-4798-8406-f753ea6ce07f · inbound

VoxSafeBench: Not Just What Is Said, but Who, How, and Where cites this paper.

VoxSafeBench: Not Just What Is Said, but Who, How, and Where TrustLLM: Trustworthiness in Large Language Models

Reference 73

Resolution
verified exact
arxiv_id, observed 2026-05-18T11:17:09.209714Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-10T10:19:28.041282Z digest=sha256:b6e22af4221b78eefd5185761f92c6aa1ec50f4798ae8cdf7e6414ec37380d7a

Observation c311ec6a-216f-4bed-9403-ac205b2e4b33 · inbound

Train Separately, Merge Together: Modular Post-Training with Mixture-of-Experts cites this paper.

Train Separately, Merge Together: Modular Post-Training with Mixture-of-Experts TrustLLM: Trustworthiness in Large Language Models

Reference 15

Resolution
verified exact
arxiv_id, observed 2026-05-18T11:17:09.209714Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-10T05:52:28.822723Z digest=sha256:1af0d297bece374ff7023880964957b35ab090fffb5feb3b12cbb4f2972bf613

Observation 5f794fa3-b6b8-44d9-aeac-e5d1386e78e2 · inbound

Vibe Medicine: Redefining Biomedical Research Through Human-AI Co-Work cites this paper.

Vibe Medicine: Redefining Biomedical Research Through Human-AI Co-Work TrustLLM: Trustworthiness in Large Language Models

Reference 39

Resolution
verified exact
arxiv_id, observed 2026-05-18T11:17:09.209714Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-08T06:16:16.121447Z digest=sha256:8a73fa700b3435c92178ef3265832a9c78ba24092b9f4a058955162ab90687af

Observation ed4aa5e3-8a1b-4790-bc99-fdc6ca734dcf · inbound

A Multi-Dimensional Audit of Politically Aligned Large Language Models cites this paper.

A Multi-Dimensional Audit of Politically Aligned Large Language Models TrustLLM: Trustworthiness in Large Language Models

Reference 19

Resolution
verified exact
arxiv_id, observed 2026-05-18T11:17:09.209714Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-08T03:38:32.761677Z digest=sha256:248873ddc0db59eaa53539a78bbde362ac3ee2315cc649d116ac8bc1109d7a2a

Observation a73c0916-974f-42cf-aaf3-ac699aed0844 · inbound

Spatiotemporal Hidden-State Dynamics as a Signature of Internal Reasoning in Large Language Models cites this paper.

Spatiotemporal Hidden-State Dynamics as a Signature of Internal Reasoning in Large Language Models TrustLLM: Trustworthiness in Large Language Models

Reference 6

Resolution
verified exact
arxiv_id, observed 2026-05-18T11:17:09.209714Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-09T17:20:19.586214Z digest=sha256:7fcaf995e234b85df280b7a5953888592a40f3e1508a6df3ddf07dc96febea38

Observation 1833038b-9211-40ef-9f83-aa074e24ec8e · inbound

Disentangling Intent from Role: Adversarial Self-Play for Persona-Invariant Safety Alignment cites this paper.

Disentangling Intent from Role: Adversarial Self-Play for Persona-Invariant Safety Alignment TrustLLM: Trustworthiness in Large Language Models

Reference 56

Resolution
verified exact
arxiv_id, observed 2026-05-18T11:17:09.209714Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-09T17:24:54.796037Z digest=sha256:e6c763336640fcd8ef4031e78e7f4adc35f8e61d2027d1bc5518d9b727144c68

Observation 01529a6c-c008-4815-ab58-555ac9f8d651 · inbound

Beyond Semantics: An Evidential Reasoning-Aware Multi-View Learning Framework for Trustworthy Mental Health Prediction cites this paper.

Beyond Semantics: An Evidential Reasoning-Aware Multi-View Learning Framework for Trustworthy Mental Health Prediction TrustLLM: Trustworthiness in Large Language Models

Reference 9

Resolution
verified exact
arxiv_id, observed 2026-05-18T11:17:09.209714Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-08T16:52:11.643353Z digest=sha256:9373078945da56cef19010812f5b981878f271584033efcba0b0e0b9749564b4

Observation 2738e14e-8fff-40a1-85f2-1ad371c5d98b · inbound

Profiling for Pennies: Unveiling the Privacy Iceberg of LLM Agents cites this paper.

Profiling for Pennies: Unveiling the Privacy Iceberg of LLM Agents TrustLLM: Trustworthiness in Large Language Models

Reference 17

Resolution
verified exact
arxiv_id, observed 2026-05-18T11:17:09.209714Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-08T09:14:04.781398Z digest=sha256:416cb2450ff0c054dd115ae2f7d36b59e5dc115cc0d54ab64e30b9a9f0e67208

Observation b416d0a9-b216-4892-8c5c-d39e671f0d7a · inbound

Trustworthy AI: Ensuring Reliability and Accountability from Models to Agents cites this paper.

Trustworthy AI: Ensuring Reliability and Accountability from Models to Agents TrustLLM: Trustworthiness in Large Language Models

Reference 80

Resolution
verified exact
arxiv_id, observed 2026-05-18T11:17:09.209714Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-12T02:21:12.882825Z digest=sha256:8269b622316301ad93499e24155c13ecdc9622f479c5048ce114b2281ba6f180

Observation 040ab9d8-731c-4c12-b822-bd33d8ef641b · inbound

Defenses at Odds: Measuring and Explaining Defense Conflicts in Large Language Models cites this paper.

Defenses at Odds: Measuring and Explaining Defense Conflicts in Large Language Models TrustLLM: Trustworthiness in Large Language Models

Reference 46

Resolution
verified exact
arxiv_id, observed 2026-05-18T11:17:09.209714Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-15T01:43:22.777232Z digest=sha256:269a562e6f76ef535e167727aae855e10fbfb9f891c20bc66eb6daf880eda6fd

Observation 48af1c22-8e1d-4eb6-b1ae-c753f5310b3f · inbound

Palette: A Modular, Controllable, and Efficient Framework for On-demand Authorized Safety Alignment Relaxation in LLMs cites this paper.

Palette: A Modular, Controllable, and Efficient Framework for On-demand Authorized Safety Alignment Relaxation in LLMs TrustLLM: Trustworthiness in Large Language Models

Reference 53

Resolution
verified exact
local_arxiv, observed 2026-06-30T16:04:52.637229Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-06-30T16:03:12.728352Z digest=sha256:3358ff64f436903c12683987dd9a55af058b714943f4af4022ef0b7b51503525

Observation ff98c025-a5de-489a-b579-2f78abf39b35 · inbound

Inform, Coach, Relate, Listen: Auditing LLM Caregiving Support Roles cites this paper.

Inform, Coach, Relate, Listen: Auditing LLM Caregiving Support Roles TrustLLM: Trustworthiness in Large Language Models

Reference 21

Resolution
verified exact
local_arxiv, observed 2026-06-29T06:03:08.884774Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-06-29T05:54:39.899517Z digest=sha256:e0bc4fb3a02443d741510187da132e1673b7141dcd8d0b9ee2a84cd6320a61cd

Observation 53d8e557-ef86-42a9-aa9c-68ff72ed1915 · inbound

Triaging Threats to Specialized Guardrails cites this paper.

Triaging Threats to Specialized Guardrails TrustLLM: Trustworthiness in Large Language Models

Reference 18

Resolution
verified exact
local_arxiv, observed 2026-06-28T22:32:44.353889Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-06-28T22:27:44.703466Z digest=sha256:ae7013f6cd7686e9e3481b8e6fa6dec1e0ff700a239aa9f14dc24d44e30fc99f

Observation 311c1276-85bf-413a-8b96-0da6fda95be7 · inbound

MESA: Improving MoE Safety Alignment via Decentralized Expertise cites this paper.

MESA: Improving MoE Safety Alignment via Decentralized Expertise TrustLLM: Trustworthiness in Large Language Models

Reference 43

Resolution
metadata mismatch
local_arxiv, observed 2026-06-28T19:52:35.287288Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-06-28T18:52:00.377915Z digest=sha256:08a19ec62ec2922de2a8eba6af325cfad046fb537c49061ef460413eb989b747

Observation eaa9f657-3040-4217-91df-7573216b88a7 · inbound

AXIOM: A Trust-First Neuro-Symbolic Execution Architecture for Verifiable Mathematical Reasoning cites this paper.

AXIOM: A Trust-First Neuro-Symbolic Execution Architecture for Verifiable Mathematical Reasoning TrustLLM: Trustworthiness in Large Language Models

Reference 7

Resolution
verified exact
local_arxiv, observed 2026-06-28T20:22:37.850598Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-06-28T18:43:39.392203Z digest=sha256:ac84277e7ed187bae975a0b521bd60e7f6f3b11682c86e89db93af77fb439c7a

Observation 58e907e1-1d81-4a74-95f3-ca8587faac83 · inbound

Testing LLM Arithmetic Reasoning Generalization with Automatic Numeric-Remapping Attacks cites this paper.

Testing LLM Arithmetic Reasoning Generalization with Automatic Numeric-Remapping Attacks TrustLLM: Trustworthiness in Large Language Models

Reference 1

Resolution
verified exact
local_arxiv, observed 2026-07-02T04:06:35.143734Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-06-28T09:27:30.923556Z digest=sha256:d5b7f28f9899b65eeb0b454aa52f048dfb8c5d4f4cef2d22abc8cdd4cecf7fad

Observation 3be491a0-c6f4-4a17-9c86-45ba674caa02 · inbound

Emergent Misalignment Can Be Induced by Sycophancy and Reversed via Alignment Gating cites this paper.

Emergent Misalignment Can Be Induced by Sycophancy and Reversed via Alignment Gating TrustLLM: Trustworthiness in Large Language Models

Reference 3

Resolution
metadata mismatch
local_arxiv, observed 2026-07-03T00:47:29.984787Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-06-27T17:03:33.199645Z digest=sha256:29e5a6e193ab3b1bedcaa600d0e58b1a0c527a84d36386b8644ee413508206b6

Observation 61519e06-4488-44e6-be40-f6985a6014d6 · inbound

When Confidence Takes the Wrong Path: Diagnosing Retrieval-State Lock-In in RAG cites this paper.

When Confidence Takes the Wrong Path: Diagnosing Retrieval-State Lock-In in RAG TrustLLM: Trustworthiness in Large Language Models

Reference 18

Resolution
verified exact
local_arxiv, observed 2026-07-04T10:09:44.105969Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-06-26T09:10:58.722292Z digest=sha256:d087d47bb2ba78f739106cdfa0ae97f08e60342e1ef5a2ae93bbdfe64407a975

Observation 92ea2e96-9ba6-4d26-8f9e-cac54e283ede · inbound

SingGuard: A Policy-Adaptive Multimodal LLM Guardrail with Dynamic Reasoning cites this paper.

SingGuard: A Policy-Adaptive Multimodal LLM Guardrail with Dynamic Reasoning TrustLLM: Trustworthiness in Large Language Models

Reference 194

Resolution
metadata mismatch
local_arxiv, observed 2026-07-04T09:59:44.781990Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-06-26T09:19:50.623741Z digest=sha256:931f38bafa48640884fea5e613eeedbf6cf8d4ed205580341cfa6dafb99a86cd

Observation c2da8abd-1716-49e7-9938-5b9653ddc045 · inbound

SingGuard: A Policy-Adaptive Multimodal LLM Guardrail with Dynamic Reasoning cites this paper.

SingGuard: A Policy-Adaptive Multimodal LLM Guardrail with Dynamic Reasoning TrustLLM: Trustworthiness in Large Language Models

Reference 193

Resolution
metadata mismatch
local_arxiv, observed 2026-07-01T18:55:59.595565Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-06-29T01:18:19.195007Z digest=sha256:b57a360425a58800d71fd8d1b87e37c07fb7d26349391212f3c83e2c65da78df

Observation f9cf52a5-332e-43d4-9b09-dfb87be3fe1d · inbound

A Red Teaming Framework for Large Language Models: A Case Study on Faithfulness Evaluation cites this paper.

A Red Teaming Framework for Large Language Models: A Case Study on Faithfulness Evaluation TrustLLM: Trustworthiness in Large Language Models

Reference 26

Resolution
metadata mismatch
local_arxiv, observed 2026-07-04T19:50:11.228177Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-06-25T20:58:53.119386Z digest=sha256:ecda31b3a0a3a9e52adc9e2f4660459558b3c01d477eee21309391253240ef03

Observation b9ccdb03-8ae3-4829-854b-b8e7bae58e4c · inbound

A Red Teaming Framework for Large Language Models: A Case Study on Faithfulness Evaluation cites this paper.

A Red Teaming Framework for Large Language Models: A Case Study on Faithfulness Evaluation TrustLLM: Trustworthiness in Large Language Models

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-02T10:16:37.400479Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T10:16:37.400479Z digest=sha256:952f7a2abf312b53864b7c06bb80486013ec5ef2afd04489fbb5f0334a8c4e34

Observation 5b0f60da-37bc-4a74-8088-f7c264729202 · inbound

When Calibration Rankings Reverse: Accuracy-Controlled Evaluation for Fair Comparison of LLMs cites this paper.

When Calibration Rankings Reverse: Accuracy-Controlled Evaluation for Fair Comparison of LLMs TrustLLM: Trustworthiness in Large Language Models

Reference 147

Resolution
metadata mismatch
local_arxiv, observed 2026-07-01T12:15:43.637641Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-07-01T02:17:19.540484Z digest=sha256:ec962115324be7023dd223b56823d36234b71b0caf6849ea9ec8f1ca32105da4

Observation 59f2273b-503a-48cd-b08d-92e58d5732df · inbound

BioTIER: A Refusal Benchmark for Targeted Biological Risk Mitigation cites this paper.

BioTIER: A Refusal Benchmark for Targeted Biological Risk Mitigation TrustLLM: Trustworthiness in Large Language Models

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-02T02:00:45.063405Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T02:00:45.063405Z digest=sha256:50b196696206b3b4c3ab9a1083f913cc55af63aa057cbad64138f5d195a6586a

Observation 35474e4f-c8c3-4c80-b3c5-21c86e0cbb99 · inbound

Fence: Specialized SLM Guardrails for LLM Applications cites this paper.

Fence: Specialized SLM Guardrails for LLM Applications TrustLLM: Trustworthiness in Large Language Models

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-02T13:21:54.308159Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T13:21:54.308159Z digest=sha256:0fa99a0a202f9966de5b17850b40ff09625832f23374b1deed0abcd57b59f293

Observation aba51b5b-1f25-4649-91a3-294785f454d0 · inbound

Reexamining zero-shot summarization: Empirical investigation of trustworthiness of LLM-summarizers cites this paper.

Reexamining zero-shot summarization: Empirical investigation of trustworthiness of LLM-summarizers TrustLLM: Trustworthiness in Large Language Models

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-01T08:46:13.605521Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T08:46:13.605521Z digest=sha256:601ef49775086def478a5f1b02a238e016dd5748915771f1479a8e9947725f76

Observation 05d7e222-abf5-44e0-affb-1291fd5dd03a · inbound

Capital Markets LLM Reliability Score (CM-LRS): From Plausible to Bankable cites this paper.

Capital Markets LLM Reliability Score (CM-LRS): From Plausible to Bankable TrustLLM: Trustworthiness in Large Language Models

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-01T07:48:35.491314Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T07:48:35.491314Z digest=sha256:7a4ce160c1336ba9f58ead896e4a18efc16157caa6b5241f6ac74968ea848905

Observation bb363c6c-7fcd-4bbb-9cf3-a5f02a5f1d1a · inbound

Cloud-Native Evaluation-as-a-Service: A Microservices Architecture for Scalable AI Monitoring with Conformal Guarantees cites this paper.

Cloud-Native Evaluation-as-a-Service: A Microservices Architecture for Scalable AI Monitoring with Conformal Guarantees TrustLLM: Trustworthiness in Large Language Models

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-02T08:50:38.795411Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T08:50:38.795411Z digest=sha256:80a2e06a683ffbf46098de68b1e39b5375439e49aa44ea49987af033a8c76950

Observation 74d15f98-5951-4c02-9903-db8f4569332f · inbound

RAG-TESTER: Automated End-to-End Testing of Retrieval-Augmented Large Language Models cites this paper.

RAG-TESTER: Automated End-to-End Testing of Retrieval-Augmented Large Language Models TrustLLM: Trustworthiness in Large Language Models

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-04T01:34:38.960505Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T01:34:38.960505Z digest=sha256:b4a6bef168c6ff3e78abf5e6ffa90ffe561b53f55aba3c7937ad2d92cabdbb0c