Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-16T04:37:48.388254Z
Paper Citation Record · LEDGER
As of 22 August 2026, this Paper Citation Record lists 100 of 100 outbound references and 1 inbound Pith citation observation for arXiv:2505.00853.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-16T04:37:48.388254Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-22T06:32:14.747728+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-05-12T04:22:06.172150Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z
100 of 100 outbound references displayed
External citation measurements
1
arxiv_reference, observed 2026-08-05T02:28:24.338817Z
Observation 399266f9-3a7d-45cd-948d-614047817435 · outbound
LLM Ethics Benchmark: A Three-Dimensional Assessment System for Evaluating Moral Reasoning in Large Language Models Concrete Problems in AI Safety
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 87a759c4-0df4-4622-9955-5f1e1028daad · outbound
LLM Ethics Benchmark: A Three-Dimensional Assessment System for Evaluating Moral Reasoning in Large Language Models Language Models as Agent Models
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 65d16a69-256e-430d-a598-18b95615c575 · outbound
LLM Ethics Benchmark: A Three-Dimensional Assessment System for Evaluating Moral Reasoning in Large Language Models Claude: A conversational ai assistant
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 226c38bb-b7cf-4314-b128-74ae045d219d · outbound
LLM Ethics Benchmark: A Three-Dimensional Assessment System for Evaluating Moral Reasoning in Large Language Models Aquino and A
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5e89fa7b-23d4-4a70-af01-9e1eef746caa · outbound
LLM Ethics Benchmark: A Three-Dimensional Assessment System for Evaluating Moral Reasoning in Large Language Models Probing pre- trained language models for cross-cultural differences in values
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 618f7f45-c46e-40d5-a34c-59cdae90624f · outbound
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e7a16ad1-da38-467e-846b-3f9a2dc65cb1 · outbound
LLM Ethics Benchmark: A Three-Dimensional Assessment System for Evaluating Moral Reasoning in Large Language Models Bandura et al
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 95217eeb-8146-4b2b-97c0-ad4c5642701d · outbound
LLM Ethics Benchmark: A Three-Dimensional Assessment System for Evaluating Moral Reasoning in Large Language Models Bender, Timnit Gebru, Angelina McMillan-Major, and Shmargaret Shmitchell
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 98db0cb4-1696-4ddf-9587-928307ef7930 · outbound
LLM Ethics Benchmark: A Three-Dimensional Assessment System for Evaluating Moral Reasoning in Large Language Models On the Opportunities and Risks of Foundation Models
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 89e09f19-12a8-421e-96f0-ad952c2cc50f · outbound
LLM Ethics Benchmark: A Three-Dimensional Assessment System for Evaluating Moral Reasoning in Large Language Models Brown, Benjamin Mann, Nick Ryder, Melanie Subbiah, Jared Kaplan, Prafulla Dhariwal, et al
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3458cb0b-d5b5-4251-bc86-7aa3ede1a61d · outbound
LLM Ethics Benchmark: A Three-Dimensional Assessment System for Evaluating Moral Reasoning in Large Language Models Unresolved cited work
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9dc6a800-24de-4508-9d68-1cfd0e5de0d9 · outbound
LLM Ethics Benchmark: A Three-Dimensional Assessment System for Evaluating Moral Reasoning in Large Language Models When Not to Trust Language Models: Investigating Effectiveness of Parametric and Non-Parametric Memories
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c64fe085-af89-40b0-80cc-12065c02b274 · outbound
LLM Ethics Benchmark: A Three-Dimensional Assessment System for Evaluating Moral Reasoning in Large Language Models Chalkidis et al
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f73bd80d-f6ff-4030-8496-6ec0b32b7d25 · outbound
LLM Ethics Benchmark: A Three-Dimensional Assessment System for Evaluating Moral Reasoning in Large Language Models Multilevel Monte Carlo methods for the Grad-Shafranov free boundary problem
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation c3ce01d4-a0f0-4d49-a4aa-9f545825b134 · outbound
LLM Ethics Benchmark: A Three-Dimensional Assessment System for Evaluating Moral Reasoning in Large Language Models The Capacity for Moral Self-Correction in Large Language Models
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ff8a65c5-ac07-4019-892f-4651bd01aa9f · outbound
LLM Ethics Benchmark: A Three-Dimensional Assessment System for Evaluating Moral Reasoning in Large Language Models Riemann zeros as quantized energies of scattering with impurities
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fbe04f05-0b35-4eab-9877-920862410467 · outbound
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3f5efe67-782d-4c47-88c2-e5d64e1725fc · outbound
LLM Ethics Benchmark: A Three-Dimensional Assessment System for Evaluating Moral Reasoning in Large Language Models Measuring moral reasoning using moral dilemmas: Evaluating reliability, validity, and differential item functioning of the behavioral defining issues test
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6511a4fa-b9de-48a8-898f-b062c1dbb8a5 · outbound
LLM Ethics Benchmark: A Three-Dimensional Assessment System for Evaluating Moral Reasoning in Large Language Models BERT: Pre-training of Deep Bidirectional Transformers for Language Understanding
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 533b32d2-b020-4233-9709-8cef4cd8c246 · outbound
LLM Ethics Benchmark: A Three-Dimensional Assessment System for Evaluating Moral Reasoning in Large Language Models Moral stories: Situated reasoning about norms, intents, actions, and their consequences
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b72cfea8-63cf-4c5a-b28a-9b014fa010e1 · outbound
LLM Ethics Benchmark: A Three-Dimensional Assessment System for Evaluating Moral Reasoning in Large Language Models Unresolved cited work
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 73bac660-f208-4fa2-9954-72cb7820f76d · outbound
LLM Ethics Benchmark: A Three-Dimensional Assessment System for Evaluating Moral Reasoning in Large Language Models MME: A Comprehensive Evaluation Benchmark for Multimodal Large Language Models
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2e4c9626-53e1-466e-9dfc-705b78ae9b6e · outbound
LLM Ethics Benchmark: A Three-Dimensional Assessment System for Evaluating Moral Reasoning in Large Language Models Artificial intelligence, values, and alignment
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0726ad03-bb68-4dcb-9d90-746e603489f5 · outbound
LLM Ethics Benchmark: A Three-Dimensional Assessment System for Evaluating Moral Reasoning in Large Language Models Graham et al
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c836b624-13d1-41e2-831c-50bd125033f2 · outbound
LLM Ethics Benchmark: A Three-Dimensional Assessment System for Evaluating Moral Reasoning in Large Language Models Moral foundations theory: The pragmatic validity of moral pluralism
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 098676a3-013e-48f4-aada-ed53d5e26154 · outbound
LLM Ethics Benchmark: A Three-Dimensional Assessment System for Evaluating Moral Reasoning in Large Language Models Unresolved cited work
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 01cb48e1-0928-4d3d-8b52-2016c8516188 · outbound
LLM Ethics Benchmark: A Three-Dimensional Assessment System for Evaluating Moral Reasoning in Large Language Models Xiezhi: An Ever-Updating Benchmark for Holistic Domain Knowledge Evaluation
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 02afc3a2-542a-466e-bdb8-70c2d3d87f0f · outbound
LLM Ethics Benchmark: A Three-Dimensional Assessment System for Evaluating Moral Reasoning in Large Language Models Unresolved cited work
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6266c894-4743-490a-b9a5-4a6585923a12 · outbound
LLM Ethics Benchmark: A Three-Dimensional Assessment System for Evaluating Moral Reasoning in Large Language Models The Righteous Mind: Why Good People are Divided by Politics and Religion
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c36c26a7-45e0-4d81-9cb7-240eece5624d · outbound
LLM Ethics Benchmark: A Three-Dimensional Assessment System for Evaluating Moral Reasoning in Large Language Models Measuring Coding Challenge Competence With APPS
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a4f5925d-3d2b-4252-9685-28a8b9d86a39 · outbound
LLM Ethics Benchmark: A Three-Dimensional Assessment System for Evaluating Moral Reasoning in Large Language Models CUAD: An Expert-Annotated NLP Dataset for Legal Contract Review
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1f06c92c-fec1-4f39-9219-3ebcf71346e7 · outbound
LLM Ethics Benchmark: A Three-Dimensional Assessment System for Evaluating Moral Reasoning in Large Language Models Measuring Mathematical Problem Solving With the MATH Dataset
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fed98156-d790-46c7-8630-b4163b6b2bbc · outbound
LLM Ethics Benchmark: A Three-Dimensional Assessment System for Evaluating Moral Reasoning in Large Language Models Enabling AI and Robotic Coaches for Physical Rehabilitation Therapy: Iterative Design and Evaluation with Therapists and Post-Stroke Survivors
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation 8c0477a4-6906-4a3a-8d92-e6ec1002c353 · outbound
LLM Ethics Benchmark: A Three-Dimensional Assessment System for Evaluating Moral Reasoning in Large Language Models Hendrycks et al
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ea736948-100c-45d0-a08b-7535d3f49880 · outbound
LLM Ethics Benchmark: A Three-Dimensional Assessment System for Evaluating Moral Reasoning in Large Language Models The weirdest people in the world? Behavioral and Brain Sciences, 33(2-3):61–83, 2010
Reference 35
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation f6ef4932-75e2-4c53-be49-70c6f4e1b198 · outbound
LLM Ethics Benchmark: A Three-Dimensional Assessment System for Evaluating Moral Reasoning in Large Language Models Unresolved cited work
Reference 36
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation c149fe60-142b-4e31-ac93-bcc51efc8880 · outbound
LLM Ethics Benchmark: A Three-Dimensional Assessment System for Evaluating Moral Reasoning in Large Language Models How is ChatGPT's behavior changing over time?
Reference 37
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation de0a7373-c8b3-483c-b8c9-81ff3124b6ec · outbound
LLM Ethics Benchmark: A Three-Dimensional Assessment System for Evaluating Moral Reasoning in Large Language Models C-Eval: A Multi-Level Multi-Discipline Chinese Evaluation Suite for Foundation Models
Reference 38
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation aa1422d1-c818-4fce-a643-5ffd0aa6bcc1 · outbound
LLM Ethics Benchmark: A Three-Dimensional Assessment System for Evaluating Moral Reasoning in Large Language Models TrustGPT: A Benchmark for Trustworthy and Responsible Large Language Models
Reference 39
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9b61b6e1-4608-40de-8070-36129ac87abb · outbound
LLM Ethics Benchmark: A Three-Dimensional Assessment System for Evaluating Moral Reasoning in Large Language Models Open-source large language models leaderboard, 2023
Reference 40
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation 6f228f24-5b15-4b50-98c4-d83ca23229ee · outbound
LLM Ethics Benchmark: A Three-Dimensional Assessment System for Evaluating Moral Reasoning in Large Language Models Inter-university Consortium for Political and Social Research, 2000
Reference 41
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation 779e205a-5f34-4e2f-8490-66f81d3d8aa7 · outbound
LLM Ethics Benchmark: A Three-Dimensional Assessment System for Evaluating Moral Reasoning in Large Language Models The global landscape of ai ethics guidelines
Reference 42
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation aff18a76-6013-49d0-8e61-17e93ba07729 · outbound
LLM Ethics Benchmark: A Three-Dimensional Assessment System for Evaluating Moral Reasoning in Large Language Models Ethics-eval: A bench- marking framework for ethical evaluation of language models, 2023
Reference 43
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation 686607ef-bf30-4bfe-91bf-592091b939a4 · outbound
LLM Ethics Benchmark: A Three-Dimensional Assessment System for Evaluating Moral Reasoning in Large Language Models Dynabench: Rethinking Benchmarking in NLP
Reference 44
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b3835249-db1e-4bda-a818-3b74ad07c320 · outbound
LLM Ethics Benchmark: A Three-Dimensional Assessment System for Evaluating Moral Reasoning in Large Language Models Unresolved cited work
Reference 45
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation 477874a3-d18c-468b-9945-4903d1be3d4e · outbound
Reference 46
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation 7d174c96-e70d-4611-8792-b58661a97943 · outbound
LLM Ethics Benchmark: A Three-Dimensional Assessment System for Evaluating Moral Reasoning in Large Language Models Koncel-Kedziorski et al
Reference 47
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation cfb384ba-44b2-4c07-a19a-0f61bd8b4ed1 · outbound
LLM Ethics Benchmark: A Three-Dimensional Assessment System for Evaluating Moral Reasoning in Large Language Models Gender, race, and intersectionality on the federal appellate bench
Reference 48
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation 9c90775c-b985-44a7-bf2a-15cfa4df7bce · outbound
LLM Ethics Benchmark: A Three-Dimensional Assessment System for Evaluating Moral Reasoning in Large Language Models SEED-Bench: Benchmarking Multimodal LLMs with Generative Comprehension
Reference 49
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e6a3457d-525c-49bf-a094-d2f542d79b25 · outbound
LLM Ethics Benchmark: A Three-Dimensional Assessment System for Evaluating Moral Reasoning in Large Language Models FullSubNet+: Channel Attention FullSubNet with Complex Spectrograms for Speech Enhancement
Reference 50
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation e45a61a7-c140-43db-aaed-07d13c08ffa0 · outbound
LLM Ethics Benchmark: A Three-Dimensional Assessment System for Evaluating Moral Reasoning in Large Language Models Logic-guided semantic representation learning for zero-shot rela- tion classification
Reference 51
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation 831e7284-0cf6-45ee-9524-d6c5dd07aa13 · outbound
LLM Ethics Benchmark: A Three-Dimensional Assessment System for Evaluating Moral Reasoning in Large Language Models Holistic Evaluation of Language Models
Reference 53
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1a65f7fe-3482-462c-9597-77996dcf39b8 · outbound
LLM Ethics Benchmark: A Three-Dimensional Assessment System for Evaluating Moral Reasoning in Large Language Models Unresolved cited work
Reference 54
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation d00db2f8-e117-4ee5-919b-2ecb525999c9 · outbound
Reference 55
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation db7db0de-16e4-4d49-baeb-e5e6ac23a2cd · outbound
LLM Ethics Benchmark: A Three-Dimensional Assessment System for Evaluating Moral Reasoning in Large Language Models MMBench: Is Your Multi-modal Model an All-around Player?
Reference 56
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a6a2dfa2-6efa-46c1-a759-b9085bd354d5 · outbound
LLM Ethics Benchmark: A Three-Dimensional Assessment System for Evaluating Moral Reasoning in Large Language Models Multi-task deep neural networks for natural language understanding
Reference 57
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation b8913c74-bd33-4f88-a351-180100d1b3f1 · outbound
LLM Ethics Benchmark: A Three-Dimensional Assessment System for Evaluating Moral Reasoning in Large Language Models Chatbot arena: Benchmarking llms in the wild with elo ratings, 2023
Reference 58
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation 616b497a-4e7b-427e-ba58-e5dc2d1649de · outbound
LLM Ethics Benchmark: A Three-Dimensional Assessment System for Evaluating Moral Reasoning in Large Language Models Unresolved cited work
Reference 59
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation c3978adb-d0a4-477d-b7ba-82ec23122cea · outbound
LLM Ethics Benchmark: A Three-Dimensional Assessment System for Evaluating Moral Reasoning in Large Language Models Algo- rithmic fairness: Choices, assumptions, and definitions
Reference 60
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation 6679eebf-8522-42c8-ac87-438bc58d6aac · outbound
LLM Ethics Benchmark: A Three-Dimensional Assessment System for Evaluating Moral Reasoning in Large Language Models Principles alone cannot guarantee ethical ai
Reference 61
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation 3b58ee3a-10f7-4a3b-bd75-2a7070fbed13 · outbound
LLM Ethics Benchmark: A Three-Dimensional Assessment System for Evaluating Moral Reasoning in Large Language Models Interpretable machine learning
Reference 62
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation a45c478c-cd8c-4499-a8f7-428a6c32bfab · outbound
LLM Ethics Benchmark: A Three-Dimensional Assessment System for Evaluating Moral Reasoning in Large Language Models Translation tutorial: 21 fairness definitions and their politics
Reference 63
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation 53fa1282-63cf-4f41-b042-16046629838f · outbound
LLM Ethics Benchmark: A Three-Dimensional Assessment System for Evaluating Moral Reasoning in Large Language Models trolley problem
Reference 64
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation ece19c91-c7d0-4f29-9527-ed7806ad51d0 · outbound
LLM Ethics Benchmark: A Three-Dimensional Assessment System for Evaluating Moral Reasoning in Large Language Models Unresolved cited work
Reference 65
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation 142dd244-394c-493f-91d0-3666184a42ff · outbound
LLM Ethics Benchmark: A Three-Dimensional Assessment System for Evaluating Moral Reasoning in Large Language Models GPT-4 Technical Report
Reference 66
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2f629139-ab87-49f5-98d7-cba9cead5372 · outbound
Reference 67
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation 3b6d4b33-e1d3-4f61-af8f-f5be3603b769 · outbound
LLM Ethics Benchmark: A Three-Dimensional Assessment System for Evaluating Moral Reasoning in Large Language Models Scaling Language Models: Methods, Analysis & Insights from Training Gopher
Reference 68
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c70c94b0-5cec-4c2f-bb9e-4e2b35ab98b0 · outbound
LLM Ethics Benchmark: A Three-Dimensional Assessment System for Evaluating Moral Reasoning in Large Language Models Closing the ai accountability gap: Defining an end-to-end framework for internal algorithmic auditing
Reference 69
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation 7b478798-fa0a-43a8-bb76-850d9afa1714 · outbound
LLM Ethics Benchmark: A Three-Dimensional Assessment System for Evaluating Moral Reasoning in Large Language Models Sentence-bert: Sentence embeddings using siamese bert- networks
Reference 70
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 55ea9473-b12c-4899-8082-2e6c2479fe9f · outbound
LLM Ethics Benchmark: A Three-Dimensional Assessment System for Evaluating Moral Reasoning in Large Language Models Unresolved cited work
Reference 71
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation 383f21b7-2b5b-4647-8f12-4142035255c9 · outbound
LLM Ethics Benchmark: A Three-Dimensional Assessment System for Evaluating Moral Reasoning in Large Language Models Adaptive testing and debugging of nlp models
Reference 72
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation 7477c323-1a21-42d5-bdf0-0b3867be9515 · outbound
LLM Ethics Benchmark: A Three-Dimensional Assessment System for Evaluating Moral Reasoning in Large Language Models Rudinger et al
Reference 73
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation 2aa00da1-5605-4f60-8d02-674d51bb3826 · outbound
Reference 74
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation 7ba73739-2ff6-4ecc-b0c0-b0707bfaded1 · outbound
Reference 75
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation 69269a41-cb9b-4dec-b4b3-51b85c4e96d0 · outbound
LLM Ethics Benchmark: A Three-Dimensional Assessment System for Evaluating Moral Reasoning in Large Language Models Singhal et al
Reference 76
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation 80cb5ae2-09d4-48d9-a794-d76c3b7cf2d6 · outbound
LLM Ethics Benchmark: A Three-Dimensional Assessment System for Evaluating Moral Reasoning in Large Language Models Open-source tools learning benchmarks, 2023
Reference 77
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation 981f48d9-e247-4ee5-a2dd-25cbf38a9821 · outbound
LLM Ethics Benchmark: A Three-Dimensional Assessment System for Evaluating Moral Reasoning in Large Language Models High-performance medicine: the convergence of human and artificial intelligence
Reference 78
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 86028ba9-313d-4213-93d7-85af4d116abe · outbound
LLM Ethics Benchmark: A Three-Dimensional Assessment System for Evaluating Moral Reasoning in Large Language Models Llama 2: Open Foundation and Fine-Tuned Chat Models
Reference 79
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b85ee886-52c3-46e8-bdc3-1d8ae7a2f650 · outbound
LLM Ethics Benchmark: A Three-Dimensional Assessment System for Evaluating Moral Reasoning in Large Language Models RoBERTweet: A BERT Language Model for Romanian Tweets
Reference 80
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation 46193ff2-c3e1-4232-bb14-30529a1e51b5 · outbound
LLM Ethics Benchmark: A Three-Dimensional Assessment System for Evaluating Moral Reasoning in Large Language Models OntoChatGPT Information System: Ontology-Driven Structured Prompts for ChatGPT Meta-Learning
Reference 81
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation a6bb601f-3bbf-48fc-8d1b-3b12079d8c9a · outbound
LLM Ethics Benchmark: A Three-Dimensional Assessment System for Evaluating Moral Reasoning in Large Language Models Neuro-symbolic Empowered Denoising Diffusion Probabilistic Models for Real-time Anomaly Detection in Industry 4.0
Reference 82
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation 88605bc5-37d3-4b52-a56c-fa08226812eb · outbound
LLM Ethics Benchmark: A Three-Dimensional Assessment System for Evaluating Moral Reasoning in Large Language Models Self-Consistency Improves Chain of Thought Reasoning in Language Models
Reference 83
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 86101bca-74bb-4827-98d6-7e582f646538 · outbound
LLM Ethics Benchmark: A Three-Dimensional Assessment System for Evaluating Moral Reasoning in Large Language Models Chain-of-Thought Prompting Elicits Reasoning in Large Language Models
Reference 84
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ce4c9d8b-e2cd-45c1-a628-514a3878dad4 · outbound
LLM Ethics Benchmark: A Three-Dimensional Assessment System for Evaluating Moral Reasoning in Large Language Models Taxonomy of risks posed by language models
Reference 85
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation 964b971e-65c4-481f-9efc-fe8047da056e · outbound
LLM Ethics Benchmark: A Three-Dimensional Assessment System for Evaluating Moral Reasoning in Large Language Models The role and limits of principles in ai ethics: towards a focus on tensions
Reference 86
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation f4ba6393-c4c0-4c73-9d30-d7cd4d07eb82 · outbound
LLM Ethics Benchmark: A Three-Dimensional Assessment System for Evaluating Moral Reasoning in Large Language Models Transformers: State- of-the-art natural language processing
Reference 87
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation aa9dca84-ffc0-4444-acd3-c68c91ed0ad3 · outbound
LLM Ethics Benchmark: A Three-Dimensional Assessment System for Evaluating Moral Reasoning in Large Language Models CValues: Measuring the Values of Chinese Large Language Models from Safety to Responsibility
Reference 88
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 191847ee-4d65-463a-88c4-f1161a1b1f10 · outbound
LLM Ethics Benchmark: A Three-Dimensional Assessment System for Evaluating Moral Reasoning in Large Language Models LVLM-eHub: A Comprehensive Evaluation Benchmark for Large Vision-Language Models
Reference 89
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 279fea2c-e9de-4ef9-884f-ffd73af4f865 · outbound
LLM Ethics Benchmark: A Three-Dimensional Assessment System for Evaluating Moral Reasoning in Large Language Models GLUE-X: Evaluating Natural Language Understanding Models from an Out-of-distribution Generalization Perspective
Reference 90
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f5792766-c14b-4fac-81cb-31b3dd1f8ac2 · outbound
LLM Ethics Benchmark: A Three-Dimensional Assessment System for Evaluating Moral Reasoning in Large Language Models Kosmos-2: Grounding Multimodal Large Language Models to the World
Reference 91
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 99624c16-5789-4d07-b095-041830f90515 · outbound
LLM Ethics Benchmark: A Three-Dimensional Assessment System for Evaluating Moral Reasoning in Large Language Models LAMM: Language-Assisted Multi-Modal Instruction-Tuning Dataset, Framework, and Benchmark
Reference 92
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 802b8577-1e88-4fd8-9f10-4a304aa202e5 · outbound
LLM Ethics Benchmark: A Three-Dimensional Assessment System for Evaluating Moral Reasoning in Large Language Models MM-Vet: Evaluating Large Multimodal Models for Integrated Capabilities
Reference 93
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 83a5518d-8949-4ec1-9a06-6176c4497ef9 · outbound
Reference 94
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation 5253eb86-bf16-4c6e-991a-f285dfb20531 · outbound
LLM Ethics Benchmark: A Three-Dimensional Assessment System for Evaluating Moral Reasoning in Large Language Models KoLA: Carefully Benchmarking World Knowledge of Large Language Models
Reference 95
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 61c6553d-e5a2-4d3a-b8af-d5bd4296f226 · outbound
LLM Ethics Benchmark: A Three-Dimensional Assessment System for Evaluating Moral Reasoning in Large Language Models Unresolved cited work
Reference 96
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation 5c8f6f66-3df2-4c73-b06c-b9c0b0d1ae7e · outbound
LLM Ethics Benchmark: A Three-Dimensional Assessment System for Evaluating Moral Reasoning in Large Language Models M3Exam: A Multilingual, Multimodal, Multilevel Benchmark for Examining Large Language Models
Reference 97
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 82e1f6db-97a6-4fc6-a326-483874931705 · outbound
LLM Ethics Benchmark: A Three-Dimensional Assessment System for Evaluating Moral Reasoning in Large Language Models Calibrate Before Use: Improving Few-Shot Performance of Language Models
Reference 98
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d2248598-2530-459f-a0d1-92e4c608773f · outbound
LLM Ethics Benchmark: A Three-Dimensional Assessment System for Evaluating Moral Reasoning in Large Language Models Link Prediction without Graph Neural Networks
Reference 99
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f1ebe5d4-d1d6-4bed-81cb-46131d2f3664 · outbound
LLM Ethics Benchmark: A Three-Dimensional Assessment System for Evaluating Moral Reasoning in Large Language Models AGIEval: A Human-Centric Benchmark for Evaluating Foundation Models
Reference 100
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b92028cd-1ee1-430a-a541-34bfd3e9fd0b · outbound
LLM Ethics Benchmark: A Three-Dimensional Assessment System for Evaluating Moral Reasoning in Large Language Models PromptRobust: Towards Evaluating the Robustness of Large Language Models on Adversarial Prompts
Reference 101
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 88f54591-d652-48b5-b21f-d120401c1208 · inbound
Agent-ValueBench: A Comprehensive Benchmark for Evaluating Agent Values LLM Ethics Benchmark: A Three-Dimensional Assessment System for Evaluating Moral Reasoning in Large Language Models
Reference 108
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.