Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-05-18T11:17:08.108565Z
Paper Citation Record · LEDGER
As of 4 August 2026, this Paper Citation Record lists 100 of 299 outbound references and 45 inbound Pith citation observations for arXiv:2401.05561.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-05-18T11:17:08.108565Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-04T06:34:03.388597+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-04T01:34:38.960505Z
A source-named dated measurement, never combined with another source.
Source: pith, observed 2026-07-04T19:50:11.226678Z
100 of 299 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation f2b7546e-1b3a-49c5-bd54-f65d529e97db · outbound
TrustLLM: Trustworthiness in Large Language Models A toolkit for text extraction and analysis for natural language processing tasks
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 6dc87d8c-cf01-4760-b178-ef5fea9cec60 · outbound
TrustLLM: Trustworthiness in Large Language Models Natural language processing: State of the art, current trends and challenges
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 52e7a414-e6a3-4e0e-9a7a-4549898e1d30 · outbound
TrustLLM: Trustworthiness in Large Language Models Wordcraft: story writing with large language models
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 0690adfa-1085-4d54-a8dc-e0568a4ab057 · outbound
TrustLLM: Trustworthiness in Large Language Models Multilingual machine translation with large language models: Empirical results and analysis
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation ded87798-abeb-4063-a612-486c2bcc813d · outbound
TrustLLM: Trustworthiness in Large Language Models https://blogs.microsoft.com/blog/2023/02/07/ reinventing-search-with-a-new-ai-powered-microsoft-bing-and-edge-your-copilot-for-the-web/
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 93998411-88da-43a2-8503-f90779080302 · outbound
TrustLLM: Trustworthiness in Large Language Models https://medium.com/whatnot-engineering/ enhancing-search-using-large-language-models-f9dcb988bdb9
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 7d233665-c017-4261-9e9b-e6cfce56cd87 · outbound
TrustLLM: Trustworthiness in Large Language Models WebGPT: Browser-assisted question-answering with human feedback
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation bff1b0b3-b48b-4511-9cca-a2e87fbf6bf6 · outbound
TrustLLM: Trustworthiness in Large Language Models https://www.projectpro.io/article/ large-language-model-use-cases-and-applications/887
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation fa095434-dcbc-49ba-98c3-4b197d7891a1 · outbound
TrustLLM: Trustworthiness in Large Language Models Code Llama: Open Foundation Models for Code
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation c669839d-7755-45d7-a266-8cc15be9f727 · outbound
TrustLLM: Trustworthiness in Large Language Models Large language models: The future of b2b software
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation e64d5080-19d8-48da-ac0a-d32862b21a2b · outbound
TrustLLM: Trustworthiness in Large Language Models Bloomberggpt: A large language model for finance
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 3ac50255-b411-4afb-b338-061655de1b55 · outbound
TrustLLM: Trustworthiness in Large Language Models Scientific discovery in the age of artificial intelligence
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation b765b235-2134-4094-b845-11eee0929722 · outbound
TrustLLM: Trustworthiness in Large Language Models Artificial Intelligence for Science in Quantum, Atomistic, and Continuum Systems
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation adb79616-6524-49ee-bc3e-d6a3e93235fb · outbound
TrustLLM: Trustworthiness in Large Language Models The impact of large language models on scientific discovery: a preliminary study using gpt-4
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation ae912b18-7b88-4276-b765-dc92f64e82e3 · outbound
TrustLLM: Trustworthiness in Large Language Models Pllama: An open-source large language model for plant science
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation f581c66b-4a58-4cc9-b75e-7d90ba06919b · outbound
TrustLLM: Trustworthiness in Large Language Models The future landscape of large language models in medicine
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 3f7610a3-606e-4da1-a98f-aa739951b492 · outbound
TrustLLM: Trustworthiness in Large Language Models ChiMed-GPT: A Chinese Medical Large Language Model with Full Training Regime and Better Alignment to Human Preferences
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 4ba49faa-1555-4be2-aae5-1e60002dcee3 · outbound
TrustLLM: Trustworthiness in Large Language Models Alpacare:instruction-tuned large language models for medical application
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation d5784197-ed8c-49c7-85a7-5df31e2b0a27 · outbound
TrustLLM: Trustworthiness in Large Language Models Davison, Quanzheng Li, Yong Chen, Hongfang Liu, and Lichao Sun
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation dbaafa24-760e-48e3-bb3e-b9664a7d4d2f · outbound
TrustLLM: Trustworthiness in Large Language Models Bianque: Balancing the questioning and suggestion ability of health llms with multi-turn health conversations polished by chatgpt
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 1182bdc9-1239-446d-8c22-a4a0457b2504 · outbound
TrustLLM: Trustworthiness in Large Language Models HuatuoGPT, towards Taming Language Model to Be a Doctor
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation afd62b14-d9b9-4151-8b19-f3a722bc3ae2 · outbound
TrustLLM: Trustworthiness in Large Language Models Chatdoctor: A medical chat model fine-tuned on a large language model meta-ai (llama) using medical domain knowledge
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 7fc96c94-95c3-4a3d-9f4a-48d6bc6dd74e · outbound
TrustLLM: Trustworthiness in Large Language Models Medicalgpt: Training medical gpt model
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 97a23ea4-43ab-4dfc-a692-b8b1e6958692 · outbound
TrustLLM: Trustworthiness in Large Language Models A domain-specific next-generation large language model (llm) or chatgpt is required for biomedical engineering and research
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 90ef08fb-b6ad-4478-8449-9fbdda9b16e8 · outbound
TrustLLM: Trustworthiness in Large Language Models Towards Generalist Biomedical AI
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 3d5de8ee-4f84-458f-8485-9ef2ba298abe · outbound
TrustLLM: Trustworthiness in Large Language Models Large language models and political science
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 3748b433-3850-406c-91fa-c24274639eb0 · outbound
TrustLLM: Trustworthiness in Large Language Models https://github.com/irlab-sdu/fuzi.mingcha
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 68037678-94d9-4c60-9292-0fe812f61ed7 · outbound
TrustLLM: Trustworthiness in Large Language Models Disc-lawllm: Fine-tuning large language models for intelligent legal services
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation e2c4e8d3-d884-4da2-ac9b-f4aa09ad1364 · outbound
TrustLLM: Trustworthiness in Large Language Models Chawla, Olaf Wiest, and Xiangliang Zhang
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation d3e938d5-ae90-4683-84d2-77f53991f4b5 · outbound
TrustLLM: Trustworthiness in Large Language Models Structured Chemistry Reasoning with Large Language Models
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation f8e95e44-d13b-45af-84e2-96bddd224863 · outbound
TrustLLM: Trustworthiness in Large Language Models Marinegpt: Unlocking secrets of "ocean" to the public
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 45553da3-af6d-4af0-b58f-2a5548b9096d · outbound
TrustLLM: Trustworthiness in Large Language Models Oceangpt: A large language model for ocean science tasks
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation c0f7efa4-8aff-444d-9dbe-6ab149e579df · outbound
TrustLLM: Trustworthiness in Large Language Models Taoli llama
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 9eb43874-decd-4dd4-a38b-96ca1dd365dd · outbound
TrustLLM: Trustworthiness in Large Language Models Artgpt-4: Artistic vision-language understanding with adapter-enhanced minigpt-4
Reference 34
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation d62334ea-3b76-4b30-9de9-17dfa9da04d2 · outbound
TrustLLM: Trustworthiness in Large Language Models Unresolved cited work
Reference 35
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 50912041-3ad1-4a01-a88d-59b30c4b3487 · outbound
TrustLLM: Trustworthiness in Large Language Models Dai, Orhan Firat, Melvin Johnson, Dmitry Lepikhin, Alexandre Passos, Sia- mak Shakeri, Emanuel Taropa, Paige Bailey, Zhifeng Chen, Eric Chu, Jonathan H
Reference 36
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 1341a0b8-52d8-4a04-b3ee-852076565bf4 · outbound
TrustLLM: Trustworthiness in Large Language Models Palm: Efficiently training massive language models
Reference 37
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 0c1a4d65-a678-4101-b51a-904b5a89482b · outbound
TrustLLM: Trustworthiness in Large Language Models How chatgpt works: A look inside large language models
Reference 38
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 98d53138-35f9-41de-8164-7891decd2683 · outbound
TrustLLM: Trustworthiness in Large Language Models LoRA: Low-Rank Adaptation of Large Language Models
Reference 39
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation d56339b2-3d20-40a6-b85a-798af7e6a869 · outbound
TrustLLM: Trustworthiness in Large Language Models QLoRA: Efficient Finetuning of Quantized LLMs
Reference 40
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation a02c6d68-7df4-4868-a16a-49aec21dd956 · outbound
TrustLLM: Trustworthiness in Large Language Models Pathways: Asynchronous distributed dataflow for ml
Reference 41
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 54bbc698-5eed-491c-933b-b877aa9833ed · outbound
TrustLLM: Trustworthiness in Large Language Models Ai alignment: A comprehensive survey
Reference 42
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 4dc02eac-8e4e-4c7e-baaa-179c0c1033a5 · outbound
TrustLLM: Trustworthiness in Large Language Models Training language models to follow instructions with human feedback
Reference 43
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 07607afd-2f0d-40dd-8792-782b8dcb8444 · outbound
TrustLLM: Trustworthiness in Large Language Models Improving Language Model Negotiation with Self-Play and In-Context Learning from AI Feedback
Reference 44
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation c9b0aad9-9dd6-45c0-ade3-e73ac5a1394c · outbound
TrustLLM: Trustworthiness in Large Language Models Principle-Driven Self-Alignment of Language Models from Scratch with Minimal Human Supervision
Reference 45
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 093f81f9-5948-41af-a9ea-7bd66ffe83fb · outbound
TrustLLM: Trustworthiness in Large Language Models Rl4f: Generating natural language feedback with reinforcement learning for repairing model outputs
Reference 46
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 8601443d-1632-4d69-907c-c0eeff835971 · outbound
TrustLLM: Trustworthiness in Large Language Models Measuring Progress on Scalable Oversight for Large Language Models
Reference 47
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 2350b892-3ea6-4f1e-9b5f-bc481c8eb5c7 · outbound
TrustLLM: Trustworthiness in Large Language Models Discovering Language Model Behaviors with Model-Written Evaluations
Reference 48
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 3b971218-5d04-47b2-94ab-5ff3d50aec18 · outbound
TrustLLM: Trustworthiness in Large Language Models Improving Factuality and Reasoning in Language Models through Multiagent Debate
Reference 49
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 1822c6ec-5d9d-41b8-aa9e-7213115c1b13 · outbound
TrustLLM: Trustworthiness in Large Language Models Characterizing Manipulation from AI Systems
Reference 50
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 841f2e03-f5ff-4121-9b72-9b295046ea61 · outbound
TrustLLM: Trustworthiness in Large Language Models RLAIF vs. RLHF: Scaling Reinforcement Learning from Human Feedback with AI Feedback
Reference 51
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation a8cf28a0-091a-425a-9a68-f21878a9f399 · outbound
TrustLLM: Trustworthiness in Large Language Models A Generalist Agent
Reference 52
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation fa881004-3ec9-458d-8ff8-1326f8aa3783 · outbound
TrustLLM: Trustworthiness in Large Language Models Constitutional AI: Harmlessness from AI Feedback
Reference 53
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation d2077b70-7dd0-4061-be6d-d022331164c4 · outbound
TrustLLM: Trustworthiness in Large Language Models The Effects of Reward Misspecification: Mapping and Mitigating Misaligned Models
Reference 54
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 3bea1e0b-b33b-462a-88fd-627d9f23d3b0 · outbound
TrustLLM: Trustworthiness in Large Language Models Cooperative inverse reinforcement learning
Reference 55
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 2f1e667a-5d98-4dcb-803c-b1a9ab60d61a · outbound
TrustLLM: Trustworthiness in Large Language Models Survey of hallucination in natural language generation
Reference 56
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 03a4e4f7-eee3-4ce1-ab81-fa6886a4a138 · outbound
TrustLLM: Trustworthiness in Large Language Models A Survey on Hallucination in Large Language Models: Principles, Taxonomy, Challenges, and Open Questions
Reference 57
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 09964268-a039-461c-a507-76ffabe403b8 · outbound
TrustLLM: Trustworthiness in Large Language Models Factuality challenges in the era of large language models
Reference 58
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation f0fe0f8b-5f0a-49bf-88e5-4a01939f063c · outbound
TrustLLM: Trustworthiness in Large Language Models Combating Misinformation in the Age of LLMs: Opportunities and Challenges
Reference 59
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation c8c59fef-2396-4038-96ac-b3eb3c3154a5 · outbound
TrustLLM: Trustworthiness in Large Language Models 10 ways cybercriminals can abuse large language models
Reference 60
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 8489c251-0396-4469-8ca0-2758d68716f8 · outbound
TrustLLM: Trustworthiness in Large Language Models Jailbroken: How Does LLM Safety Training Fail?
Reference 61
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 4a2094e0-2a25-4150-8d07-36e430454598 · outbound
TrustLLM: Trustworthiness in Large Language Models Unraveling the link between translations and gender bias in llms
Reference 62
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 24436732-944f-4a9e-9718-1066e8b8c3a5 · outbound
TrustLLM: Trustworthiness in Large Language Models Navigating the biases in llm generative ai: A guide to responsible implementation
Reference 63
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 7562b62f-f0be-4df5-99d8-006de124854d · outbound
TrustLLM: Trustworthiness in Large Language Models Large language models may leak personal data
Reference 64
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 225241b2-9bf9-4567-89e7-280534c4f6ef · outbound
TrustLLM: Trustworthiness in Large Language Models Deid-gpt: Zero-shot medical text de-identification by gpt-4
Reference 65
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 4a987fff-772a-4891-9db9-28be8bf80211 · outbound
TrustLLM: Trustworthiness in Large Language Models What does it mean to align ai with human values?
Reference 66
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 54305edd-5d14-460b-8ee5-a232b9131252 · outbound
Reference 67
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 1d3f8eda-4cf1-49cc-87aa-dfcba66f9987 · outbound
TrustLLM: Trustworthiness in Large Language Models Ai at meta
Reference 68
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation f2604d1a-0cf0-41bc-929c-a2f50db9d375 · outbound
TrustLLM: Trustworthiness in Large Language Models Llama 2: Open Foundation and Fine-Tuned Chat Models
Reference 69
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation dacd4082-fada-42c9-bcd6-eccb45ff3db4 · outbound
TrustLLM: Trustworthiness in Large Language Models Holistic Evaluation of Language Models
Reference 70
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 7133d9f0-4fd0-4606-bbf8-076a4c12371c · outbound
TrustLLM: Trustworthiness in Large Language Models DecodingTrust: A Comprehensive Assessment of Trustworthiness in GPT Models
Reference 71
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 5c4d3f45-cc05-44bc-8fc8-e190eb8f5f1b · outbound
TrustLLM: Trustworthiness in Large Language Models Trustworthy LLMs: a Survey and Guideline for Evaluating Large Language Models' Alignment
Reference 72
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation f80178cc-aec5-402b-b6c5-a92bbc41ba05 · outbound
TrustLLM: Trustworthiness in Large Language Models Do-Not-Answer: A Dataset for Evaluating Safeguards in LLMs
Reference 73
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 9dd37d76-2f6c-4b01-8e03-63adfa5b6950 · outbound
TrustLLM: Trustworthiness in Large Language Models Chatbot arena leaderboard week 8: Introducing mt-bench and vicuna-33b
Reference 74
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 13bdfdcf-5d0a-4771-b1e6-a3b37657acb4 · outbound
TrustLLM: Trustworthiness in Large Language Models The big benchmarks collection - a open-llm-leaderboard collection
Reference 75
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation fcbfa731-8aa7-40e6-bc1b-4b1e1dc45dcd · outbound
TrustLLM: Trustworthiness in Large Language Models https://platform.openai.com/docs/guides/moderation
Reference 76
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 9ed9d9d5-4405-4513-8741-6ae38f48e506 · outbound
TrustLLM: Trustworthiness in Large Language Models The foundation model transparency index
Reference 77
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation c4509149-f6c5-49f1-936c-cfd39c17430b · outbound
TrustLLM: Trustworthiness in Large Language Models Unresolved cited work
Reference 78
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 3597cca9-b5aa-4aeb-af41-4871c3096910 · outbound
TrustLLM: Trustworthiness in Large Language Models Ernie - baidu yiyan
Reference 79
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation e90a6ad9-c56f-4b54-bf1a-dcc3730ea6ce · outbound
TrustLLM: Trustworthiness in Large Language Models Attention is all you need
Reference 80
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 41418d48-616d-453c-ab8e-8c92adccc940 · outbound
TrustLLM: Trustworthiness in Large Language Models an open assistant for everyone by laion
Reference 81
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 76ebbf8d-a854-40ba-a1dd-f20d9205b14b · outbound
TrustLLM: Trustworthiness in Large Language Models Gonzalez, Ion Stoica, and Eric P
Reference 82
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 6c35a320-bdd0-45f5-a6da-950e1bdc8298 · outbound
TrustLLM: Trustworthiness in Large Language Models https://nvlpubs.nist.gov/nistpubs/ ai/NIST.AI.100-1.pdf
Reference 83
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 99e38ce1-b67b-4fb4-83ac-ed546aa32f7c · outbound
TrustLLM: Trustworthiness in Large Language Models Enron email dataset
Reference 84
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation acc153dc-42c9-48d5-96b3-e0226b5d6543 · outbound
TrustLLM: Trustworthiness in Large Language Models Guest editors’ introduction: machine ethics
Reference 85
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation be040ea0-e057-4426-9cf3-06fa61806efe · outbound
TrustLLM: Trustworthiness in Large Language Models Machine ethics: Creating an ethical intelligent agent
Reference 86
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 02025677-03fd-4f7c-859c-9c8b52f261cf · outbound
TrustLLM: Trustworthiness in Large Language Models Emergent Abilities of Large Language Models
Reference 87
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 7ae3047f-8d4d-42e6-bb11-203122fbb27c · outbound
TrustLLM: Trustworthiness in Large Language Models Chain-of-thought prompting elicits reasoning in large language models
Reference 88
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation bba74d8d-1ca0-4110-aac3-e8bd91586c95 · outbound
TrustLLM: Trustworthiness in Large Language Models Scaling Instruction-Finetuned Language Models
Reference 89
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation bcb2691c-cb97-4b6d-8b9e-21ac98881482 · outbound
TrustLLM: Trustworthiness in Large Language Models Exploring the limits of transfer learning with a unified text-to-text transformer
Reference 90
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation fa1c89e3-76a9-47a9-9493-e4290ccf391b · outbound
TrustLLM: Trustworthiness in Large Language Models Scaling Laws for Neural Language Models
Reference 91
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 576152b0-6f81-4d4b-97a0-89d797b42952 · outbound
TrustLLM: Trustworthiness in Large Language Models Training Compute-Optimal Large Language Models
Reference 92
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 441eeb77-e4c6-4f1d-8066-b312a393e33b · outbound
TrustLLM: Trustworthiness in Large Language Models Proximal Policy Optimization Algorithms
Reference 93
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 42c92d63-24f4-4771-ad82-0777d5589509 · outbound
TrustLLM: Trustworthiness in Large Language Models RAFT: Reward rAnked FineTuning for Generative Foundation Model Alignment
Reference 94
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 45b8b23c-2054-42df-bf3e-bfb40f708d26 · outbound
TrustLLM: Trustworthiness in Large Language Models OpenChat: Advancing Open-source Language Models with Mixed-Quality Data
Reference 95
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 51f31bb9-f395-427e-8558-ebef725cee31 · outbound
TrustLLM: Trustworthiness in Large Language Models Chain of Hindsight Aligns Language Models with Feedback
Reference 96
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 037599fd-e5a9-47d5-9b37-633391134655 · outbound
TrustLLM: Trustworthiness in Large Language Models Training Socially Aligned Language Models on Simulated Social Interactions
Reference 97
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation c1f558d4-6c9c-445e-b5c6-483ad92e0747 · outbound
TrustLLM: Trustworthiness in Large Language Models Yu, Qiang Yang, and Xing Xie
Reference 98
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation a368994f-6309-4724-903f-16d32c1729ee · outbound
TrustLLM: Trustworthiness in Large Language Models Can chatgpt forecast stock price movements? return pre- dictability and large language models
Reference 99
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 370d8f83-950f-40dd-9c0f-e5224a579b7e · outbound
TrustLLM: Trustworthiness in Large Language Models Sentiment analysis in the era of large language models: A reality check
Reference 100
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation e5944dee-3fd8-4fc3-9260-08aec534976f · inbound
Large Language Models: A Survey TrustLLM: Trustworthiness in Large Language Models
Reference 221
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 8853d8da-4b81-4463-9835-b3e09969d32c · inbound
JailbreakBench: An Open Robustness Benchmark for Jailbreaking Large Language Models TrustLLM: Trustworthiness in Large Language Models
Reference 49
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 7de9155a-012d-4aa1-8dcb-1ebcc44a60de · inbound
Trustworthiness in Retrieval-Augmented Generation Systems: A Survey TrustLLM: Trustworthiness in Large Language Models
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 46c76b80-b86d-4d6d-b0ab-a93d47e301be · inbound
Entry-level guide to the use of large language models for medical research TrustLLM: Trustworthiness in Large Language Models
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation bd754c48-a272-4f69-83ef-dd0af4ff3f1b · inbound
Opportunities and Challenges of Large Language Models for Low-Resource Languages in Humanities Research TrustLLM: Trustworthiness in Large Language Models
Reference 58
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 2dd88203-ff68-4191-a56e-9eea7459501c · inbound
Enabling Global, Human-Centered Explanations for LLMs:From Tokens to Interpretable Code and Test Generation TrustLLM: Trustworthiness in Large Language Models
Reference 60
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation f306397b-db9e-45bd-bd57-23b43b64ac12 · inbound
ReGA: Model-Based Safeguard for LLMs via Representation-Guided Abstraction TrustLLM: Trustworthiness in Large Language Models
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation fdea7240-6eda-4d2a-9986-f2f0fe0dec68 · inbound
Downgrade to Upgrade: Optimizer Simplification Enhances Robustness in LLM Unlearning TrustLLM: Trustworthiness in Large Language Models
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation d7032d39-2799-4dd7-aa96-b2a198868bcf · inbound
Erase to Improve: Erasable Reinforcement Learning for Search-Augmented LLMs TrustLLM: Trustworthiness in Large Language Models
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation bbf26d3a-a66b-4ed6-87db-0fc1853522d2 · inbound
Leak@$k$: Unlearning Does Not Make LLMs Forget Under Probabilistic Decoding TrustLLM: Trustworthiness in Large Language Models
Reference 38
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e752d3d0-1aee-489e-b3ed-8273c6d81d6f · inbound
OutSafe-Bench: A Benchmark for Multimodal Offensive Content Detection in Large Language Models TrustLLM: Trustworthiness in Large Language Models
Reference 51
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation ad293aa8-23ec-4169-aedc-2ca13de390d0 · inbound
On the Factual Consistency of Text-based Explainable Recommendation Models TrustLLM: Trustworthiness in Large Language Models
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 4bb73bbe-4165-45a5-aa9b-f8ec8177015d · inbound
SafeCRS: Personalized Safety Alignment for LLM-Based Conversational Recommender Systems TrustLLM: Trustworthiness in Large Language Models
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5b64b74b-f1f4-46a0-a682-64ca777b4076 · inbound
Shorter, but Still Trustworthy? An Empirical Study of Chain-of-Thought Compression TrustLLM: Trustworthiness in Large Language Models
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 0e3ea996-6cc6-4b19-bfba-bb2e154434d3 · inbound
Guardian-as-an-Advisor: Advancing Next-Generation Guardian Models for Trustworthy LLMs TrustLLM: Trustworthiness in Large Language Models
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation d5b68b55-d08d-430d-9674-b90a026ffb27 · inbound
AVID: A Benchmark for Omni-Modal Audio-Visual Inconsistency Understanding via Agent-Driven Construction TrustLLM: Trustworthiness in Large Language Models
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 21b38b88-2462-4798-8406-f753ea6ce07f · inbound
VoxSafeBench: Not Just What Is Said, but Who, How, and Where TrustLLM: Trustworthiness in Large Language Models
Reference 73
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation c311ec6a-216f-4bed-9403-ac205b2e4b33 · inbound
Train Separately, Merge Together: Modular Post-Training with Mixture-of-Experts TrustLLM: Trustworthiness in Large Language Models
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 5f794fa3-b6b8-44d9-aeac-e5d1386e78e2 · inbound
Vibe Medicine: Redefining Biomedical Research Through Human-AI Co-Work TrustLLM: Trustworthiness in Large Language Models
Reference 39
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation ed4aa5e3-8a1b-4790-bc99-fdc6ca734dcf · inbound
A Multi-Dimensional Audit of Politically Aligned Large Language Models TrustLLM: Trustworthiness in Large Language Models
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation a73c0916-974f-42cf-aaf3-ac699aed0844 · inbound
Spatiotemporal Hidden-State Dynamics as a Signature of Internal Reasoning in Large Language Models TrustLLM: Trustworthiness in Large Language Models
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 1833038b-9211-40ef-9f83-aa074e24ec8e · inbound
Disentangling Intent from Role: Adversarial Self-Play for Persona-Invariant Safety Alignment TrustLLM: Trustworthiness in Large Language Models
Reference 56
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 01529a6c-c008-4815-ab58-555ac9f8d651 · inbound
Beyond Semantics: An Evidential Reasoning-Aware Multi-View Learning Framework for Trustworthy Mental Health Prediction TrustLLM: Trustworthiness in Large Language Models
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 2738e14e-8fff-40a1-85f2-1ad371c5d98b · inbound
Profiling for Pennies: Unveiling the Privacy Iceberg of LLM Agents TrustLLM: Trustworthiness in Large Language Models
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation b416d0a9-b216-4892-8c5c-d39e671f0d7a · inbound
Trustworthy AI: Ensuring Reliability and Accountability from Models to Agents TrustLLM: Trustworthiness in Large Language Models
Reference 80
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 040ab9d8-731c-4c12-b822-bd33d8ef641b · inbound
Defenses at Odds: Measuring and Explaining Defense Conflicts in Large Language Models TrustLLM: Trustworthiness in Large Language Models
Reference 46
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 48af1c22-8e1d-4eb6-b1ae-c753f5310b3f · inbound
Palette: A Modular, Controllable, and Efficient Framework for On-demand Authorized Safety Alignment Relaxation in LLMs TrustLLM: Trustworthiness in Large Language Models
Reference 53
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation ff98c025-a5de-489a-b579-2f78abf39b35 · inbound
Inform, Coach, Relate, Listen: Auditing LLM Caregiving Support Roles TrustLLM: Trustworthiness in Large Language Models
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 53d8e557-ef86-42a9-aa9c-68ff72ed1915 · inbound
Triaging Threats to Specialized Guardrails TrustLLM: Trustworthiness in Large Language Models
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 311c1276-85bf-413a-8b96-0da6fda95be7 · inbound
MESA: Improving MoE Safety Alignment via Decentralized Expertise TrustLLM: Trustworthiness in Large Language Models
Reference 43
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation eaa9f657-3040-4217-91df-7573216b88a7 · inbound
AXIOM: A Trust-First Neuro-Symbolic Execution Architecture for Verifiable Mathematical Reasoning TrustLLM: Trustworthiness in Large Language Models
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 58e907e1-1d81-4a74-95f3-ca8587faac83 · inbound
Testing LLM Arithmetic Reasoning Generalization with Automatic Numeric-Remapping Attacks TrustLLM: Trustworthiness in Large Language Models
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 3be491a0-c6f4-4a17-9c86-45ba674caa02 · inbound
Emergent Misalignment Can Be Induced by Sycophancy and Reversed via Alignment Gating TrustLLM: Trustworthiness in Large Language Models
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 61519e06-4488-44e6-be40-f6985a6014d6 · inbound
When Confidence Takes the Wrong Path: Diagnosing Retrieval-State Lock-In in RAG TrustLLM: Trustworthiness in Large Language Models
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 92ea2e96-9ba6-4d26-8f9e-cac54e283ede · inbound
SingGuard: A Policy-Adaptive Multimodal LLM Guardrail with Dynamic Reasoning TrustLLM: Trustworthiness in Large Language Models
Reference 194
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation c2da8abd-1716-49e7-9938-5b9653ddc045 · inbound
SingGuard: A Policy-Adaptive Multimodal LLM Guardrail with Dynamic Reasoning TrustLLM: Trustworthiness in Large Language Models
Reference 193
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation f9cf52a5-332e-43d4-9b09-dfb87be3fe1d · inbound
A Red Teaming Framework for Large Language Models: A Case Study on Faithfulness Evaluation TrustLLM: Trustworthiness in Large Language Models
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation b9ccdb03-8ae3-4829-854b-b8e7bae58e4c · inbound
A Red Teaming Framework for Large Language Models: A Case Study on Faithfulness Evaluation TrustLLM: Trustworthiness in Large Language Models
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5b0f60da-37bc-4a74-8088-f7c264729202 · inbound
When Calibration Rankings Reverse: Accuracy-Controlled Evaluation for Fair Comparison of LLMs TrustLLM: Trustworthiness in Large Language Models
Reference 147
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 59f2273b-503a-48cd-b08d-92e58d5732df · inbound
BioTIER: A Refusal Benchmark for Targeted Biological Risk Mitigation TrustLLM: Trustworthiness in Large Language Models
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 35474e4f-c8c3-4c80-b3c5-21c86e0cbb99 · inbound
Fence: Specialized SLM Guardrails for LLM Applications TrustLLM: Trustworthiness in Large Language Models
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation aba51b5b-1f25-4649-91a3-294785f454d0 · inbound
Reexamining zero-shot summarization: Empirical investigation of trustworthiness of LLM-summarizers TrustLLM: Trustworthiness in Large Language Models
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 05d7e222-abf5-44e0-affb-1291fd5dd03a · inbound
Capital Markets LLM Reliability Score (CM-LRS): From Plausible to Bankable TrustLLM: Trustworthiness in Large Language Models
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bb363c6c-7fcd-4bbb-9cf3-a5f02a5f1d1a · inbound
Cloud-Native Evaluation-as-a-Service: A Microservices Architecture for Scalable AI Monitoring with Conformal Guarantees TrustLLM: Trustworthiness in Large Language Models
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 74d15f98-5951-4c02-9903-db8f4569332f · inbound
RAG-TESTER: Automated End-to-End Testing of Retrieval-Augmented Large Language Models TrustLLM: Trustworthiness in Large Language Models
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.