Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-07T00:39:42.008599Z
Paper Citation Record · LEDGER
As of 7 August 2026, this Paper Citation Record lists 100 of 205 outbound references and 2 inbound Pith citation observations for arXiv:2506.12958.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-07T00:39:42.008599Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-07-04T00:30:00.449270Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-07-04T00:39:16.515159Z
100 of 205 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 22e77fb1-afc9-447b-ad28-c2aa00bf703f · outbound
Domain Specific Benchmarks for Evaluating Multimodal Large Language Models GPT-4 Technical Report
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2a72d66f-913d-4913-b4fd-347fd990f34d · outbound
Domain Specific Benchmarks for Evaluating Multimodal Large Language Models Gemini: A Family of Highly Capable Multimodal Models
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fc12aaa7-ac0a-4015-9df2-edbde6fc0274 · outbound
Domain Specific Benchmarks for Evaluating Multimodal Large Language Models On the Opportunities and Risks of Foundation Models
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2713e6d7-e1c5-4ab0-8c91-1be92280ff06 · outbound
Domain Specific Benchmarks for Evaluating Multimodal Large Language Models The Llama 3 Herd of Models
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation dab422af-0405-4654-a183-5bb34a8b6071 · outbound
Domain Specific Benchmarks for Evaluating Multimodal Large Language Models Phi-4 Technical Report
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2d5ec7c9-4c18-4cc9-ac07-9ff8ff0158c0 · outbound
Domain Specific Benchmarks for Evaluating Multimodal Large Language Models Islam, A
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2e1fe199-bb2d-41af-97c4-c0f839cd60d5 · outbound
Domain Specific Benchmarks for Evaluating Multimodal Large Language Models Top General Performance = Top Domain Performance? DomainCodeBench: A Multi-domain Code Generation Benchmark
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 12671620-04e8-46cb-b642-1c477a68f677 · outbound
Domain Specific Benchmarks for Evaluating Multimodal Large Language Models MMRo: Are Multimodal LLMs Eligible as the Brain for In-Home Robotics?
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 896f59b5-69ca-47b5-a9eb-80e6d851d4c7 · outbound
Domain Specific Benchmarks for Evaluating Multimodal Large Language Models Unresolved cited work
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e5c6c6a6-9ccf-49b2-9f0d-8b1301c63d1e · outbound
Domain Specific Benchmarks for Evaluating Multimodal Large Language Models Kernan Freire, C
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation dd1fd5f1-5b1b-4920-ae13-1bda3c35e56d · outbound
Domain Specific Benchmarks for Evaluating Multimodal Large Language Models Unresolved cited work
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 73e3fb07-81fa-4fd3-b35a-81b054ecfbed · outbound
Domain Specific Benchmarks for Evaluating Multimodal Large Language Models FDM-Bench: A Comprehensive Benchmark for Evaluating Large Language Models in Additive Manufacturing Tasks
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9fdfcfdb-258b-47a6-8e08-62b3f0b30309 · outbound
Domain Specific Benchmarks for Evaluating Multimodal Large Language Models Fakih, R
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 25ca9dca-1ffa-4aff-809b-27353d0b0d36 · outbound
Domain Specific Benchmarks for Evaluating Multimodal Large Language Models Tizaoui, R
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5a0e13cb-75f2-42cd-bfc2-6b2335e1a9d8 · outbound
Domain Specific Benchmarks for Evaluating Multimodal Large Language Models Incorporating Large Language Models into Production Systems for Enhanced Task Automation and Flexibility
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 05d52643-c9e0-42da-b3e9-9c86aff446f2 · outbound
Domain Specific Benchmarks for Evaluating Multimodal Large Language Models Ogundare, S
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b9dafaaa-0960-46f5-9b17-3d146a08ea74 · outbound
Domain Specific Benchmarks for Evaluating Multimodal Large Language Models Unresolved cited work
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7cb939d5-6d1e-4d3a-af26-a83d0ec9f842 · outbound
Domain Specific Benchmarks for Evaluating Multimodal Large Language Models Large Language Models for Supply Chain Optimization
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 63640395-7108-44df-8b52-f090b168807f · outbound
Domain Specific Benchmarks for Evaluating Multimodal Large Language Models Raman, A
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b424281b-a063-4232-949d-841fb688cb6f · outbound
Domain Specific Benchmarks for Evaluating Multimodal Large Language Models Developing a Scalable Benchmark for Assessing Large Language Models in Knowledge Graph Engineering
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0c13e9a4-b54e-4bbd-97a8-2a3e11e63f50 · outbound
Domain Specific Benchmarks for Evaluating Multimodal Large Language Models bench authors, Beyond the imitation game: Quantify- ing and extrapolating the capabilities of language models, Transactions on Machine Learning Research (2023)
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 47aa9c45-576d-4b16-8f9c-02610589c765 · outbound
Domain Specific Benchmarks for Evaluating Multimodal Large Language Models Tracking the Moving Target: A Framework for Continuous Evaluation of LLM Test Generation in Industry
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 45969fe8-4c72-4500-9457-8a31e174c01a · outbound
Domain Specific Benchmarks for Evaluating Multimodal Large Language Models Unresolved cited work
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7b743598-c7ce-4dea-8175-0ada8e673ef5 · outbound
Domain Specific Benchmarks for Evaluating Multimodal Large Language Models A Survey on Large Language Models for Software Engineering
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cc39fc4f-2268-4700-b4e3-773a03027744 · outbound
Domain Specific Benchmarks for Evaluating Multimodal Large Language Models Unresolved cited work
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7ee961ce-1b53-4b12-a41e-1fa960769104 · outbound
Domain Specific Benchmarks for Evaluating Multimodal Large Language Models Unresolved cited work
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d35071e2-6a71-4045-afaa-bd7f04821575 · outbound
Domain Specific Benchmarks for Evaluating Multimodal Large Language Models Do Large Language Model Benchmarks Test Reliability?
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 17853490-1b3d-4947-a2fe-55801a832645 · outbound
Domain Specific Benchmarks for Evaluating Multimodal Large Language Models Trustworthy LLMs: a Survey and Guideline for Evaluating Large Language Models' Alignment
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bb7f8507-1658-4704-845c-110ed6a0b86d · outbound
Domain Specific Benchmarks for Evaluating Multimodal Large Language Models TEOChat: A Large Vision-Language Assistant for Temporal Earth Observation Data
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 17eb8b0d-4c3d-43df-a02b-1b633efb8d05 · outbound
Domain Specific Benchmarks for Evaluating Multimodal Large Language Models Xiong, F
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6d95f1cf-d097-4862-953f-7717116ca134 · outbound
Domain Specific Benchmarks for Evaluating Multimodal Large Language Models Good at captioning, bad at counting: Benchmarking GPT-4V on Earth observation data
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation eeca5cda-ba16-4d8f-a648-34cd90fe1baf · outbound
Domain Specific Benchmarks for Evaluating Multimodal Large Language Models STBench: Assessing the Ability of Large Language Models in Spatio-Temporal Analysis
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5b9f2c1f-ef36-4b4d-b51d-6a480b1277d6 · outbound
Domain Specific Benchmarks for Evaluating Multimodal Large Language Models Sapkota, R
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cf07693c-673d-4c7d-b64e-d7774d47fbc0 · outbound
Domain Specific Benchmarks for Evaluating Multimodal Large Language Models GEOBench-VLM: Benchmarking Vision-Language Models for Geospatial Tasks
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 015b86be-6aba-4bcd-b163-d57f470b86b4 · outbound
Domain Specific Benchmarks for Evaluating Multimodal Large Language Models Roberts, T
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 34c328bb-c4b8-44cc-a63f-6325c4738d5e · outbound
Domain Specific Benchmarks for Evaluating Multimodal Large Language Models RSUniVLM: A Unified Vision Language Model for Remote Sensing via Granularity-oriented Mixture of Experts
Reference 36
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0ae8ff29-4cde-422d-9dc3-fd203fb620dc · outbound
Domain Specific Benchmarks for Evaluating Multimodal Large Language Models INS-MMBench: A Comprehensive Benchmark for Evaluating LVLMs' Performance in Insurance
Reference 37
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fd27debd-ccab-4ec9-92b2-ad2e80910fd5 · outbound
Domain Specific Benchmarks for Evaluating Multimodal Large Language Models Measuring Massive Multitask Language Understanding
Reference 38
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ac4b8cd4-8a23-424e-9d88-6719ffc816ec · outbound
Domain Specific Benchmarks for Evaluating Multimodal Large Language Models Unresolved cited work
Reference 39
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bfc5a169-50ce-4d1c-a3ec-7ea13cd751fe · outbound
Domain Specific Benchmarks for Evaluating Multimodal Large Language Models IsoBench: Benchmarking Multimodal Foundation Models on Isomorphic Representations
Reference 40
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation baa06e3f-ace0-4e5d-a9bd-4ce2120fd3b3 · outbound
Domain Specific Benchmarks for Evaluating Multimodal Large Language Models VisScience: An Extensive Benchmark for Evaluating K12 Educational Multi-modal Scientific Reasoning
Reference 41
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b036bad6-e0c1-4035-8597-db4a9a169700 · outbound
Domain Specific Benchmarks for Evaluating Multimodal Large Language Models Unresolved cited work
Reference 42
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 662fbc80-6ba8-4a53-b9e9-fa7625348af0 · outbound
Domain Specific Benchmarks for Evaluating Multimodal Large Language Models Anand, J
Reference 43
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 459e5491-99e7-4447-84c4-e24d25bb03ee · outbound
Domain Specific Benchmarks for Evaluating Multimodal Large Language Models Are large language models superhuman chemists?
Reference 44
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 76e129ff-fae7-4056-b42c-d2b49e42a482 · outbound
Domain Specific Benchmarks for Evaluating Multimodal Large Language Models Unresolved cited work
Reference 45
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 163d8c6b-3890-4bb6-996f-6c7ca2aeb36c · outbound
Domain Specific Benchmarks for Evaluating Multimodal Large Language Models Unresolved cited work
Reference 46
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0575fa89-83cc-4a9d-93c6-9b3f9879ba57 · outbound
Domain Specific Benchmarks for Evaluating Multimodal Large Language Models LlaSMol: Advancing Large Language Models for Chemistry with a Large-Scale, Comprehensive, High-Quality Instruction Tuning Dataset
Reference 47
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2f3217dd-187c-45ef-a60f-244cc6f64d7f · outbound
Domain Specific Benchmarks for Evaluating Multimodal Large Language Models MaScQA: A Question Answering Dataset for Investigating Materials Science Knowledge of Large Language Models
Reference 48
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 876bb9b6-fd8a-4277-b1b4-f02d497effad · outbound
Domain Specific Benchmarks for Evaluating Multimodal Large Language Models LLM4Mat-Bench: Benchmarking Large Language Models for Materials Property Prediction
Reference 49
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 406b8ac6-39e1-4b40-b5df-853c56e65809 · outbound
Domain Specific Benchmarks for Evaluating Multimodal Large Language Models PRESTO: Progressive Pretraining Enhances Synthetic Chemistry Outcomes
Reference 50
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d6541dfb-3b14-4440-bd58-ac9be06d78d7 · outbound
Domain Specific Benchmarks for Evaluating Multimodal Large Language Models DrugLLM: Open Large Language Model for Few-shot Molecule Generation
Reference 51
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3eb95384-25ce-4589-b119-8b09891e2168 · outbound
Domain Specific Benchmarks for Evaluating Multimodal Large Language Models Unresolved cited work
Reference 52
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 415390c2-1227-465c-a149-c8999ed55dd9 · outbound
Domain Specific Benchmarks for Evaluating Multimodal Large Language Models Unresolved cited work
Reference 53
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 93e70b19-314a-48b5-acf5-924ffa0a41da · outbound
Domain Specific Benchmarks for Evaluating Multimodal Large Language Models WeatherQA: Can Multimodal Language Models Reason about Severe Weather?
Reference 54
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 41ddf636-1a72-4f42-96eb-c37e8d8960f4 · outbound
Domain Specific Benchmarks for Evaluating Multimodal Large Language Models CLLMate: A Multimodal Benchmark for Weather and Climate Events Forecasting
Reference 55
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 486c1bb9-f746-4797-84f8-7a75078d10f1 · outbound
Domain Specific Benchmarks for Evaluating Multimodal Large Language Models Unresolved cited work
Reference 56
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3f31b45e-c978-44a9-b2c7-21cf96fea40b · outbound
Domain Specific Benchmarks for Evaluating Multimodal Large Language Models DisasterQA: A Benchmark for Assessing the performance of LLMs in Disaster Response
Reference 57
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2cc6b699-4760-4877-b558-46206ff6317c · outbound
Domain Specific Benchmarks for Evaluating Multimodal Large Language Models VayuBuddy: an LLM-Powered Chatbot to Democratize Air Quality Insights
Reference 58
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6c9f8696-69cb-47b7-a1e1-b90cfe7cac76 · outbound
Domain Specific Benchmarks for Evaluating Multimodal Large Language Models Pafilis, S
Reference 59
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9eb06b2f-9f5f-4d75-85f6-9c591c69a61d · outbound
Domain Specific Benchmarks for Evaluating Multimodal Large Language Models Abdelmageed, F
Reference 60
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8b5a2ea4-e28b-4999-bdba-2db041a08b62 · outbound
Domain Specific Benchmarks for Evaluating Multimodal Large Language Models Are We Done with MMLU?
Reference 61
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2152909a-75c5-43c6-a38a-7605a0e9bed3 · outbound
Domain Specific Benchmarks for Evaluating Multimodal Large Language Models Unresolved cited work
Reference 62
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 69fbe33a-b4be-4c9f-805d-4e5daa08b57d · outbound
Domain Specific Benchmarks for Evaluating Multimodal Large Language Models Unresolved cited work
Reference 63
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation dcd02d8a-d1de-4d28-aacf-f6280cf63d25 · outbound
Domain Specific Benchmarks for Evaluating Multimodal Large Language Models Unresolved cited work
Reference 64
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8df08264-9de2-4a25-9421-78599b826a8a · outbound
Domain Specific Benchmarks for Evaluating Multimodal Large Language Models Edwards, C
Reference 65
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c85c5585-2d02-4e62-83a8-08718fe9be40 · outbound
Domain Specific Benchmarks for Evaluating Multimodal Large Language Models Unresolved cited work
Reference 66
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 29ea0a6b-e898-411e-ad3f-02daf9765111 · outbound
Domain Specific Benchmarks for Evaluating Multimodal Large Language Models Davies, M
Reference 67
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d09c6ae9-ac48-4e4e-a5fc-94a52fb674a1 · outbound
Domain Specific Benchmarks for Evaluating Multimodal Large Language Models Unresolved cited work
Reference 68
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f36fd6d9-a808-4022-af84-409d5deabdbc · outbound
Domain Specific Benchmarks for Evaluating Multimodal Large Language Models ClimateLLM: Efficient Weather Forecasting via Frequency-Aware Large Language Models
Reference 69
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fcbc7896-1894-40fe-b9dd-46279ab48d66 · outbound
Domain Specific Benchmarks for Evaluating Multimodal Large Language Models Sachdeva, N
Reference 70
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 583c3fae-d05a-4577-b23f-b0ebfb43c860 · outbound
Domain Specific Benchmarks for Evaluating Multimodal Large Language Models Unresolved cited work
Reference 71
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 87d2ce94-80e9-4f9f-84f4-36e3d86f93aa · outbound
Domain Specific Benchmarks for Evaluating Multimodal Large Language Models Unresolved cited work
Reference 72
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 10caa8a0-72a9-4ec0-98a5-e370804464e0 · outbound
Domain Specific Benchmarks for Evaluating Multimodal Large Language Models Cambrian-1: A Fully Open, Vision-Centric Exploration of Multimodal LLMs
Reference 73
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ec9c7d18-c642-40cc-8285-665c74826ec4 · outbound
Domain Specific Benchmarks for Evaluating Multimodal Large Language Models Unresolved cited work
Reference 74
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5455bd5c-2f4d-43bb-a23f-b15c5e66c49b · outbound
Domain Specific Benchmarks for Evaluating Multimodal Large Language Models Unresolved cited work
Reference 75
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation baec2d01-64f0-431e-9c3b-1387ed1e70d7 · outbound
Domain Specific Benchmarks for Evaluating Multimodal Large Language Models Unresolved cited work
Reference 76
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 0caf4c69-bfbb-473c-a4b0-3963e0ca5858 · outbound
Domain Specific Benchmarks for Evaluating Multimodal Large Language Models Malla, C
Reference 77
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c9162ea3-2b41-4700-ba06-e3dfa2ae021f · outbound
Domain Specific Benchmarks for Evaluating Multimodal Large Language Models Unresolved cited work
Reference 78
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 63c4f861-85b8-434c-87bf-7574c59db5c1 · outbound
Domain Specific Benchmarks for Evaluating Multimodal Large Language Models Unresolved cited work
Reference 79
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c4ab97bc-27da-4f8e-9709-5ffd9cff91f4 · outbound
Domain Specific Benchmarks for Evaluating Multimodal Large Language Models Unresolved cited work
Reference 80
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 13e19500-827b-4e73-9199-aba13927bfa1 · outbound
Domain Specific Benchmarks for Evaluating Multimodal Large Language Models Language Prompt for Autonomous Driving
Reference 81
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation dd094b8a-70e8-4055-978a-0dea2ce54259 · outbound
Domain Specific Benchmarks for Evaluating Multimodal Large Language Models Logic Meets Magic: LLMs Cracking Smart Contract Vulnerabilities
Reference 82
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 996de77b-5907-42c1-86da-fd6bac78b1a6 · outbound
Domain Specific Benchmarks for Evaluating Multimodal Large Language Models LLM-SmartAudit: Advanced Smart Contract Vulnerability Detection
Reference 83
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7dd2f9f2-e621-41c1-a09b-05ad1de09f48 · outbound
Domain Specific Benchmarks for Evaluating Multimodal Large Language Models ACFIX: Guiding LLMs with Mined Common RBAC Practices for Context-Aware Repair of Access Control Vulnerabilities in Smart Contracts
Reference 84
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1a3b0867-df9e-4ef7-9dcf-32fde5867a49 · outbound
Domain Specific Benchmarks for Evaluating Multimodal Large Language Models Unresolved cited work
Reference 85
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2c5e8517-ba74-46e8-a590-7cff7c738c7b · outbound
Domain Specific Benchmarks for Evaluating Multimodal Large Language Models Ethereum Price Prediction Employing Large Language Models for Short-term and Few-shot Forecasting
Reference 86
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation e6879f06-fd77-45b5-a5d3-6705c5f77814 · outbound
Domain Specific Benchmarks for Evaluating Multimodal Large Language Models Exploring LLM Cryptocurrency Trading Through Fact-Subjectivity Aware Reasoning
Reference 87
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bb59d99e-d90b-4048-864e-43c08836a429 · outbound
Domain Specific Benchmarks for Evaluating Multimodal Large Language Models Large Language Models for Blockchain Security: A Systematic Literature Review
Reference 88
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ff40ad15-0fd6-4ffb-b399-4d85b0dd2199 · outbound
Domain Specific Benchmarks for Evaluating Multimodal Large Language Models Large Language Models in Cryptocurrency Securities Cases: Can a GPT Model Meaningfully Assist Lawyers?
Reference 89
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 63d28df8-52ec-428a-852e-e42117a0f360 · outbound
Domain Specific Benchmarks for Evaluating Multimodal Large Language Models Unresolved cited work
Reference 90
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 454c4815-e6e7-4b6b-a6d9-fcc3d40cc7c8 · outbound
Domain Specific Benchmarks for Evaluating Multimodal Large Language Models Unresolved cited work
Reference 91
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7d63b3f8-0a63-489c-8ce9-aebedf8a93a2 · outbound
Domain Specific Benchmarks for Evaluating Multimodal Large Language Models Unresolved cited work
Reference 92
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 756b38bf-bb3d-4f4e-bd0e-7aed0526441d · outbound
Domain Specific Benchmarks for Evaluating Multimodal Large Language Models Unresolved cited work
Reference 93
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ce57edaf-9138-4a1a-9e8a-22fb1fcea8b4 · outbound
Domain Specific Benchmarks for Evaluating Multimodal Large Language Models WHEN FLUE MEETS FLANG: Benchmarks and Large Pre-trained Language Model for Financial Domain
Reference 94
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3cbfd683-b20e-4fe3-9865-c3d859638272 · outbound
Domain Specific Benchmarks for Evaluating Multimodal Large Language Models LogLLM: Log-based Anomaly Detection Using Large Language Models
Reference 95
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9b129a3a-eb73-4630-ac44-4989b94424b5 · outbound
Domain Specific Benchmarks for Evaluating Multimodal Large Language Models Real-Time Anomaly Detection and Reactive Planning with Large Language Models
Reference 96
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 83ec3f38-6365-431b-9bab-001ec26a5cc6 · outbound
Domain Specific Benchmarks for Evaluating Multimodal Large Language Models Zhang, D
Reference 97
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ff885e08-d07b-475e-be47-43864ae90972 · outbound
Domain Specific Benchmarks for Evaluating Multimodal Large Language Models Unresolved cited work
Reference 98
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 215976d9-09fb-4dac-9b96-68245e69699a · outbound
Domain Specific Benchmarks for Evaluating Multimodal Large Language Models Unresolved cited work
Reference 99
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d3d98513-624c-4268-8633-5f40cf866eda · outbound
Domain Specific Benchmarks for Evaluating Multimodal Large Language Models MathChat: Benchmarking Mathematical Reasoning and Instruction Following in Multi-Turn Interactions
Reference 100
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 203daee5-81f4-4de5-b65e-c0cebffff883 · inbound
ResearchClawBench: A Benchmark for End-to-End Autonomous Scientific Research Domain Specific Benchmarks for Evaluating Multimodal Large Language Models
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation b41655b5-201e-41b0-9551-43758eaa1c12 · inbound
ResearchClawBench: A Benchmark for End-to-End Autonomous Scientific Research Domain Specific Benchmarks for Evaluating Multimodal Large Language Models
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.