Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-07T12:35:33.032518Z
Paper Citation Record · LEDGER
As of 7 August 2026, this Paper Citation Record lists 30 of 30 outbound references and 0 inbound Pith citation observations for arXiv:2505.24635.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-07T12:35:33.032518Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links
A source-named dated measurement, never combined with another source.
Source: cited_works
30 of 30 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation b584db00-d352-45d2-8f3e-d42e8cb31f84 · outbound
Disentangling Language and Culture for Evaluating Multilingual Large Language Models Unresolved cited work
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fe36fe63-bf95-4b02-aba1-fe168ef7da96 · outbound
Disentangling Language and Culture for Evaluating Multilingual Large Language Models Toward Generalizable Evaluation in the LLM Era: A Survey Beyond Benchmarks
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 545853b1-e495-4d69-a88b-671cfb2c1227 · outbound
Disentangling Language and Culture for Evaluating Multilingual Large Language Models Model Utility Law: Evaluating LLMs beyond Performance through Mechanism Interpretable Metric
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d6e1a612-e347-49a3-b5e8-66b55159d98f · outbound
Disentangling Language and Culture for Evaluating Multilingual Large Language Models CulturalBench: A Robust, Diverse, and Challenging Cultural Benchmark by Human-AI CulturalTeaming
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0f093d86-4add-44d5-b7bc-6130c22afe9c · outbound
Disentangling Language and Culture for Evaluating Multilingual Large Language Models Think you have Solved Question Answering? Try ARC, the AI2 Reasoning Challenge
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e5a57296-a27f-417e-84a0-0d78ab4a73cf · outbound
Disentangling Language and Culture for Evaluating Multilingual Large Language Models Unresolved cited work
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation ab413b52-840a-4d12-bef7-4b465ede556a · outbound
Disentangling Language and Culture for Evaluating Multilingual Large Language Models Unresolved cited work
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 97426cd6-c677-4952-b9da-1efedfbd4f00 · outbound
Disentangling Language and Culture for Evaluating Multilingual Large Language Models The Llama 3 Herd of Models
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5cc99fa0-6ce5-4cf3-bf32-eaa45ada50ad · outbound
Disentangling Language and Culture for Evaluating Multilingual Large Language Models Unresolved cited work
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4c34f7a2-6192-4306-b09f-bd0a425a6909 · outbound
Disentangling Language and Culture for Evaluating Multilingual Large Language Models Intrinsic Test of Unlearning Using Parametric Knowledge Traces
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a4467368-c0c4-45b9-9801-02a593adf7b1 · outbound
Disentangling Language and Culture for Evaluating Multilingual Large Language Models Unresolved cited work
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation a11b7aff-1b05-483f-b640-8515b7c6c8e7 · outbound
Disentangling Language and Culture for Evaluating Multilingual Large Language Models BHASA: A Holistic Southeast Asian Linguistic and Cultural Evaluation Suite for Large Language Models
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation db639b68-38c1-4584-87dd-41e5125fb336 · outbound
Disentangling Language and Culture for Evaluating Multilingual Large Language Models SeaExam and SeaBench: Benchmarking LLMs with Local Multilingual Questions in Southeast Asia
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 417313c1-cc50-42f6-b690-1c93f4b9ea54 · outbound
Disentangling Language and Culture for Evaluating Multilingual Large Language Models Crosslingual Generalization through Multitask Finetuning
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7a09cff2-1343-4168-b394-1863624cde60 · outbound
Disentangling Language and Culture for Evaluating Multilingual Large Language Models Unresolved cited work
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation b6627754-7a28-409a-a819-fe97787bc1b0 · outbound
Disentangling Language and Culture for Evaluating Multilingual Large Language Models Unresolved cited work
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5e829491-d525-4fe6-9ff4-02dc8b388965 · outbound
Disentangling Language and Culture for Evaluating Multilingual Large Language Models Unresolved cited work
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 0d910d80-1073-4ae0-a48c-5777ee2eb970 · outbound
Disentangling Language and Culture for Evaluating Multilingual Large Language Models GPT-4 Technical Report
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 274dd890-6f2b-4fe5-abfa-88d2fc3bf5c4 · outbound
Disentangling Language and Culture for Evaluating Multilingual Large Language Models Qwen2.5 Technical Report
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1740e2e0-9e3d-4055-8641-d8456358748a · outbound
Disentangling Language and Culture for Evaluating Multilingual Large Language Models Language Models are Multilingual Chain-of-Thought Reasoners
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3e978c99-7ad8-4c14-b83a-5dc943c536eb · outbound
Disentangling Language and Culture for Evaluating Multilingual Large Language Models Language-Specific Neurons: The Key to Multilingual Capabilities in Large Language Models
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9e974ea5-6772-40b4-af04-873cf1c17ffc · outbound
Disentangling Language and Culture for Evaluating Multilingual Large Language Models Gemma 2: Improving Open Language Models at a Practical Size
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 248d82b0-cc93-41e4-b373-26ab7dfe7199 · outbound
Disentangling Language and Culture for Evaluating Multilingual Large Language Models Unresolved cited work
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 34fba750-ab18-4cf6-8e74-9eccf946f6b6 · outbound
Disentangling Language and Culture for Evaluating Multilingual Large Language Models Unresolved cited work
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2d4d9e63-387f-4f55-9ef2-9ed1152d6fc5 · outbound
Disentangling Language and Culture for Evaluating Multilingual Large Language Models Do Llamas Work in English? On the Latent Language of Multilingual Transformers
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation eb7bcc25-5d00-407a-95a0-8b0f2b563f27 · outbound
Disentangling Language and Culture for Evaluating Multilingual Large Language Models Unresolved cited work
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5dd1b83d-7f22-4f67-b534-26abce526a3c · outbound
Disentangling Language and Culture for Evaluating Multilingual Large Language Models Unresolved cited work
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 6e5cdfa0-728f-4af4-bc80-471c884a1b90 · outbound
Disentangling Language and Culture for Evaluating Multilingual Large Language Models SeaLLMs 3: Open Foundation and Chat Multilingual Large Language Models for Southeast Asian Languages
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 11197da8-c143-4207-adcd-a3ab214724d4 · outbound
Disentangling Language and Culture for Evaluating Multilingual Large Language Models How do Large Language Models Handle Multilingualism?
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ae484c60-081a-429b-bece-63d9e534e71d · outbound
Disentangling Language and Culture for Evaluating Multilingual Large Language Models Judging LLM-as-a-Judge with MT-Bench and Chatbot Arena
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
No inbound Pith citation observations are available.