Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
Paper Citation Record · LEDGER
As of 7 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 51 inbound Pith citation observations for arXiv:2311.16452.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-07T05:19:33.202015Z
A source-named dated measurement, never combined with another source.
Source: pith, observed 2026-08-05T02:28:24.338817Z
0 of 0 outbound references displayed
External citation measurements
169
pith, observed 2026-08-05T02:28:24.338817Z
No outbound reference observations are available for this paper version.
Observation fd1fe60a-68e2-4196-a587-1c877d8a65b3 · inbound
Can an LLM Learn Preferences from Choice Data? Can Generalist Foundation Models Outcompete Special-Purpose Tuning? Case Study in Medicine
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 40cf7b2b-c48c-4b16-bb3f-232de0c584ce · inbound
Capabilities of Gemini Models in Medicine Can Generalist Foundation Models Outcompete Special-Purpose Tuning? Case Study in Medicine
Reference 179
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation a519c327-9818-40e0-b5f5-d09500c33bca · inbound
AgentClinic: a multimodal agent benchmark to evaluate AI in simulated clinical environments Can Generalist Foundation Models Outcompete Special-Purpose Tuning? Case Study in Medicine
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 009b931c-2c08-40bb-a714-12ac3ad3bc94 · inbound
TextGrad: Automatic "Differentiation" via Text Can Generalist Foundation Models Outcompete Special-Purpose Tuning? Case Study in Medicine
Reference 66
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 4b101104-0530-4786-99c8-f4f0c95200ff · inbound
GPT-4o System Card Can Generalist Foundation Models Outcompete Special-Purpose Tuning? Case Study in Medicine
Reference 40
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 2c9d2596-c233-4a81-91f0-a9f73bb0d42f · inbound
HuatuoGPT-o1, Towards Medical Complex Reasoning with LLMs Can Generalist Foundation Models Outcompete Special-Purpose Tuning? Case Study in Medicine
Reference 47
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 590815fc-94c2-46c3-b16f-481ce68d607f · inbound
Ask Patients with Patience: Enabling LLMs for Human-Centric Medical Dialogue with Grounded Reasoning Can Generalist Foundation Models Outcompete Special-Purpose Tuning? Case Study in Medicine
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation f8520b19-0c50-45f4-94db-ffbdc7d9dc38 · inbound
Advances and Challenges in Foundation Agents: From Brain-Inspired Intelligence to Evolutionary, Collaborative, and Safe Systems Can Generalist Foundation Models Outcompete Special-Purpose Tuning? Case Study in Medicine
Reference 183
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 3cd2eaad-b15f-42f2-bf48-c9c9468ef675 · inbound
Evaluating LLMs Across Multi-Cognitive Levels: From Medical Knowledge Mastery to Scenario-Based Problem Solving Can Generalist Foundation Models Outcompete Special-Purpose Tuning? Case Study in Medicine
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c6a7b146-9077-46f3-ac73-8b87b1a0b477 · inbound
MasHost Builds It All: Autonomous Multi-Agent System Directed by Reinforcement Learning Can Generalist Foundation Models Outcompete Special-Purpose Tuning? Case Study in Medicine
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d5bf19ba-eb2d-412d-a731-3b80054b4a6a · inbound
Towards Effective Complementary Security Analysis using Large Language Models Can Generalist Foundation Models Outcompete Special-Purpose Tuning? Case Study in Medicine
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e6cd44f3-eb7c-4740-8564-9cdf7933fda2 · inbound
Knowledge Augmented Finetuning Matters in both RAG and Agent Based Dialog Systems Can Generalist Foundation Models Outcompete Special-Purpose Tuning? Case Study in Medicine
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 11300047-0100-41d6-8187-d94d83fae0bd · inbound
Dissecting Clinical Reasoning in Language Models: A Comparative Study of Prompts and Model Adaptation Strategies Can Generalist Foundation Models Outcompete Special-Purpose Tuning? Case Study in Medicine
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation da4f5f8d-d53a-454c-811c-5521ea9f913b · inbound
ALIGN: Prompt-based Attribute Alignment for Reliable, Responsible, and Personalized LLM-based Decision-Making Can Generalist Foundation Models Outcompete Special-Purpose Tuning? Case Study in Medicine
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9850c560-7d20-4d05-8d5e-bf9bf6e0f933 · inbound
HIVMedQA: Benchmarking large language models for HIV medical decision support Can Generalist Foundation Models Outcompete Special-Purpose Tuning? Case Study in Medicine
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ec171c18-daf7-4b0b-8ba4-81621992f1ce · inbound
Making Prompts First-Class Citizens for Adaptive LLM Pipelines Can Generalist Foundation Models Outcompete Special-Purpose Tuning? Case Study in Medicine
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 32e84c90-d21c-415e-9b3a-d8d4390ea1a3 · inbound
The Non-Determinism of Small LLMs: Evidence of Low Answer Consistency in Repetition Trials of Standard Multiple-Choice Benchmarks Can Generalist Foundation Models Outcompete Special-Purpose Tuning? Case Study in Medicine
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 797b5eb6-1460-48b0-b778-10461e3117c3 · inbound
Testing for LLM response differences: the case of a composite null consisting of semantically irrelevant query perturbations Can Generalist Foundation Models Outcompete Special-Purpose Tuning? Case Study in Medicine
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 59938cd2-eb3b-481c-9f91-ce3c1fd65b46 · inbound
Teaching large language models to reason like expert diagnosticians Can Generalist Foundation Models Outcompete Special-Purpose Tuning? Case Study in Medicine
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a36f5af5-f04a-407c-a62f-048d90e1344c · inbound
Measuring Competency, Not Performance: Item-Aware Evaluation Across Medical Benchmarks Can Generalist Foundation Models Outcompete Special-Purpose Tuning? Case Study in Medicine
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation acac3c1a-a841-4656-8c0a-1cf74bfbb339 · inbound
Toward a Vision-Language Foundation Model for Medical Data: Multimodal Dataset and Benchmarks for Vietnamese PET/CT Report Generation Can Generalist Foundation Models Outcompete Special-Purpose Tuning? Case Study in Medicine
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ca6da547-36a8-400f-8806-751ac54a6d42 · inbound
Cross-Platform Evaluation of Large Language Model Safety in Pediatric Consultations: Evolution of Adversarial Robustness and the Scale Paradox Can Generalist Foundation Models Outcompete Special-Purpose Tuning? Case Study in Medicine
Reference 41
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 29678fc4-f2d5-4061-beee-70719ec01730 · inbound
Group Selection as a Safeguard Against AI Substitution Can Generalist Foundation Models Outcompete Special-Purpose Tuning? Case Study in Medicine
Reference 3440
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3dd5afd2-3ac7-426d-96df-443e2c8b4ed8 · inbound
Med-V1: Small Language Models for Zero-shot and Scalable Biomedical Evidence Attribution Can Generalist Foundation Models Outcompete Special-Purpose Tuning? Case Study in Medicine
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation cee22403-0557-4751-bce5-60d25e451921 · inbound
Medical Reasoning with Large Language Models: A Survey and MR-Bench Can Generalist Foundation Models Outcompete Special-Purpose Tuning? Case Study in Medicine
Reference 71
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation cf618542-ba1f-4615-af0c-2cef29eebcfd · inbound
Evaluating Small Open LLMs for Medical Question Answering: A Practical Framework Can Generalist Foundation Models Outcompete Special-Purpose Tuning? Case Study in Medicine
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 170f340b-0b20-49c0-867d-b02514f1934b · inbound
MADE: A Living Benchmark for Multi-Label Text Classification with Uncertainty Quantification of Medical Device Adverse Events Can Generalist Foundation Models Outcompete Special-Purpose Tuning? Case Study in Medicine
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 3dfccda0-0ba3-4105-aaee-76073463d261 · inbound
SymptomAI: Toward a Conversational AI Agent for Everyday Symptom Assessment Can Generalist Foundation Models Outcompete Special-Purpose Tuning? Case Study in Medicine
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation c4e771f2-2391-4501-9761-faeee07a3991 · inbound
SymptomAI: Toward a Conversational AI Agent for Everyday Symptom Assessment Can Generalist Foundation Models Outcompete Special-Purpose Tuning? Case Study in Medicine
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 956e374c-8716-464b-93d1-484ce7d98ddd · inbound
Decodable but Not Corrected by Fixed Residual-Stream Linear Steering: Evidence from Medical LLM Failure Regimes Can Generalist Foundation Models Outcompete Special-Purpose Tuning? Case Study in Medicine
Reference 51
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation b770c0a4-751a-40e3-ad07-ae5342690514 · inbound
AgentSlimming: Towards Efficient and Cost-Aware Multi-Agent Systems Can Generalist Foundation Models Outcompete Special-Purpose Tuning? Case Study in Medicine
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation ec767918-802c-497d-a451-f65b1ab606e5 · inbound
AgentRx: A Benchmark Study of LLM Agents for Multimodal Clinical Prediction Tasks Can Generalist Foundation Models Outcompete Special-Purpose Tuning? Case Study in Medicine
Reference 45
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation ed19ff2b-9d15-4ee3-8e36-240d3f797754 · inbound
AgentRx: A Benchmark Study of LLM Agents for Multimodal Clinical Prediction Tasks Can Generalist Foundation Models Outcompete Special-Purpose Tuning? Case Study in Medicine
Reference 117
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation d9ca0eea-1b14-4e43-95a8-113184e51f75 · inbound
PrivScope: Task-scoped Disclosure Control for Hybrid Agentic Systems Can Generalist Foundation Models Outcompete Special-Purpose Tuning? Case Study in Medicine
Reference 41
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation e7d11e7c-7a1a-4370-b91e-59d5084f1092 · inbound
AgentCo-op: Retrieval-Based Synthesis of Interoperable Multi-Agent Workflows Can Generalist Foundation Models Outcompete Special-Purpose Tuning? Case Study in Medicine
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 3677f857-2551-467d-8867-a9b0271e0562 · inbound
NeuroQA: A Large-Scale Image-Grounded Benchmark for 3D Brain MRI Understanding Can Generalist Foundation Models Outcompete Special-Purpose Tuning? Case Study in Medicine
Reference 44
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation d8b688e7-c21b-4f51-bcb4-ea292fe8dc8f · inbound
A Clinically Validated Foundation Model for Comprehensive Lung Pathology Interpretation Can Generalist Foundation Models Outcompete Special-Purpose Tuning? Case Study in Medicine
Reference 45
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 2ade2156-6d2f-4fea-89c3-df43243b4886 · inbound
A Clinically Validated Foundation Model for Comprehensive Lung Pathology Interpretation Can Generalist Foundation Models Outcompete Special-Purpose Tuning? Case Study in Medicine
Reference 59
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e7045d5b-d427-4604-8f75-4b2dd5c3f277 · inbound
SURGENT: A Surgical Multi-Agent Assistance System Across the Perioperative Workflow Can Generalist Foundation Models Outcompete Special-Purpose Tuning? Case Study in Medicine
Reference 37
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 29bac44a-75a6-472b-b992-b5726d270afa · inbound
FAM-Bench: A Multimodal Benchmark for Condition-Aware Food-as-Medicine Reasoning Can Generalist Foundation Models Outcompete Special-Purpose Tuning? Case Study in Medicine
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 81fec4d2-6705-4bd2-8830-503868681192 · inbound
Search-Time Contamination in Deep Research Agents: Measuring Performance Inflation in Public Benchmark Evaluation Can Generalist Foundation Models Outcompete Special-Purpose Tuning? Case Study in Medicine
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 98cd100f-046e-4cab-8e8d-401729f7e5c9 · inbound
Towards Unified and Data-Efficient Prognostics and Health Management with Tabular Foundation Models Can Generalist Foundation Models Outcompete Special-Purpose Tuning? Case Study in Medicine
Reference 43
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 6fcbd93a-2fe8-4267-a2a0-dd2e2d0d2d03 · inbound
Small LLMs for Biomedical Claim Verification: Cost-Effective Fine-Tuning, Structural Dataset Shortcuts, and Cross-Domain Generalization Can Generalist Foundation Models Outcompete Special-Purpose Tuning? Case Study in Medicine
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation cc07f78c-1766-4260-ba4f-012bf3996a37 · inbound
When Medical Safety Alignment Fails: A Benchmark for Evaluating LLMs on High-Risk Medical Queries Can Generalist Foundation Models Outcompete Special-Purpose Tuning? Case Study in Medicine
Reference 49
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 76eaf936-269e-4997-9b52-d97e308b3a72 · inbound
MedEvoEval: Evaluating Continual Evolution of Doctor Agents through Simulated Clinical Episodes Can Generalist Foundation Models Outcompete Special-Purpose Tuning? Case Study in Medicine
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation e23af757-cfea-49b7-a47a-ff1b4e5124e0 · inbound
FaithMed: Training LLMs For Faithful Evidence-Based Medical Reasoning Can Generalist Foundation Models Outcompete Special-Purpose Tuning? Case Study in Medicine
Reference 65
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 292b8c9a-bbfd-4729-81a0-eeeca18aea4f · inbound
Toward Trustworthy Large Language Model Agents in Healthcare Can Generalist Foundation Models Outcompete Special-Purpose Tuning? Case Study in Medicine
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 66436146-ad99-4c73-8af2-4dc13a4f617a · inbound
LLM Agents for Deliberative Collaboration: A Study on Joint Decision Making Under Partial Observability Can Generalist Foundation Models Outcompete Special-Purpose Tuning? Case Study in Medicine
Reference 74
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation aab801bd-2527-413d-9b33-3157153fa581 · inbound
In-Context Learning for Wound Classification with Small Multimodal Language Models Can Generalist Foundation Models Outcompete Special-Purpose Tuning? Case Study in Medicine
Reference 53
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ea9593bb-f474-4e0e-9d9d-568a23ab9391 · inbound
DeepLens Diagnosis Agent: Agentic Workflow Design Lets a Small Reasoning Model Compete with Frontier LLMs Can Generalist Foundation Models Outcompete Special-Purpose Tuning? Case Study in Medicine
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 39c08e43-e7f2-4c98-bf2b-6245d031a19a · inbound
PatientAgentBench: A Benchmark Framework for Evaluating Patient-Facing Health AI Agents Can Generalist Foundation Models Outcompete Special-Purpose Tuning? Case Study in Medicine
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.