Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-07T14:27:50.052728Z
Paper Citation Record · LEDGER
As of 7 August 2026, this Paper Citation Record lists 39 of 39 outbound references and 1 inbound Pith citation observation for arXiv:2505.18831.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-07T14:27:50.052728Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-07T05:20:04.591495Z
A source-named dated measurement, never combined with another source.
Source: pith, observed 2026-08-07T05:20:04.993495Z
39 of 39 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 275b4e78-d049-43e7-9e7d-499a0839ab6b · outbound
Enhancing LLMs' Reasoning-Intensive Multimedia Search Capabilities through Fine-Tuning and Reinforcement Learning Unresolved cited work
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 031634fd-eb23-4a55-9843-b3d924c0a3cd · outbound
Enhancing LLMs' Reasoning-Intensive Multimedia Search Capabilities through Fine-Tuning and Reinforcement Learning Adapting Knowledge Prompt Tuning for Enhanced Automated Program Repair
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 64778890-b6a6-4654-a09a-bf2c576f6b6b · outbound
Enhancing LLMs' Reasoning-Intensive Multimedia Search Capabilities through Fine-Tuning and Reinforcement Learning MARS: a Multimodal Alignment and Ranking System for Few-Shot Segmentation
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 26421d0a-b553-4bbe-8545-ea9fa2490670 · outbound
Enhancing LLMs' Reasoning-Intensive Multimedia Search Capabilities through Fine-Tuning and Reinforcement Learning Unresolved cited work
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8b14969c-8e5d-4b38-a3a8-68e785e093ae · outbound
Enhancing LLMs' Reasoning-Intensive Multimedia Search Capabilities through Fine-Tuning and Reinforcement Learning Deep reinforcement learning from human preferences
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a1385966-e092-4749-9504-d1e5f63e5406 · outbound
Enhancing LLMs' Reasoning-Intensive Multimedia Search Capabilities through Fine-Tuning and Reinforcement Learning Neural Spacetimes for DAG Representation Learning
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation e9ccada1-68ee-43c4-ba63-fc0410daa3ce · outbound
Enhancing LLMs' Reasoning-Intensive Multimedia Search Capabilities through Fine-Tuning and Reinforcement Learning DeepSeek-V2: A Strong, Economical, and Efficient Mixture-of-Experts Language Model
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e589d879-6482-434a-914b-44d8abc6423e · outbound
Enhancing LLMs' Reasoning-Intensive Multimedia Search Capabilities through Fine-Tuning and Reinforcement Learning MM-IFEngine: Towards Multimodal Instruction Following
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d80d2367-28ca-4496-964d-aa9e971b0b74 · outbound
Enhancing LLMs' Reasoning-Intensive Multimedia Search Capabilities through Fine-Tuning and Reinforcement Learning Specializing Smaller Language Models towards Multi-Step Reasoning
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cdaeaf76-f5ee-4ca1-9485-b665ec98a63e · outbound
Enhancing LLMs' Reasoning-Intensive Multimedia Search Capabilities through Fine-Tuning and Reinforcement Learning Benchmarking Multimodal CoT Reward Model Stepwise by Visual Program
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 392c4abf-df88-4a78-b058-ab2240a985fb · outbound
Enhancing LLMs' Reasoning-Intensive Multimedia Search Capabilities through Fine-Tuning and Reinforcement Learning Unresolved cited work
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 91f6d3af-5d5b-4245-abe8-cda51feb0c9d · outbound
Enhancing LLMs' Reasoning-Intensive Multimedia Search Capabilities through Fine-Tuning and Reinforcement Learning Supervised Contrastive Learning for Pre-trained Language Model Fine-tuning
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ea22e792-a2e9-4167-834f-3ea8266a9e64 · outbound
Enhancing LLMs' Reasoning-Intensive Multimedia Search Capabilities through Fine-Tuning and Reinforcement Learning LLM4GNAS: A Large Language Model Based Toolkit for Graph Neural Architecture Search
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation de8244f9-100b-431b-b4c8-1a2cba1335e1 · outbound
Enhancing LLMs' Reasoning-Intensive Multimedia Search Capabilities through Fine-Tuning and Reinforcement Learning MMSearch: Benchmarking the Potential of Large Models as Multi-modal Search Engines
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1b27107d-9ac8-4be0-88f6-40b1ce23ec4e · outbound
Enhancing LLMs' Reasoning-Intensive Multimedia Search Capabilities through Fine-Tuning and Reinforcement Learning FinSphere, a Real-Time Stock Analysis Agent Powered by Instruction-Tuned LLMs and Domain Tools
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ae46c352-c78f-49c3-8cff-7b10dd8c232d · outbound
Enhancing LLMs' Reasoning-Intensive Multimedia Search Capabilities through Fine-Tuning and Reinforcement Learning Search-R1: Training LLMs to Reason and Leverage Search Engines with Reinforcement Learning
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fa6067af-3fdc-4d38-b2e0-94fb346c2aa7 · outbound
Enhancing LLMs' Reasoning-Intensive Multimedia Search Capabilities through Fine-Tuning and Reinforcement Learning Unresolved cited work
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 22e71aaf-666b-4a64-87f1-da82be82236f · outbound
Enhancing LLMs' Reasoning-Intensive Multimedia Search Capabilities through Fine-Tuning and Reinforcement Learning RLAIF vs. RLHF: Scaling Reinforcement Learning from Human Feedback with AI Feedback
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9b5d2bbe-8e6a-4cff-85f2-8db749303448 · outbound
Enhancing LLMs' Reasoning-Intensive Multimedia Search Capabilities through Fine-Tuning and Reinforcement Learning The Impact of Reasoning Step Length on Large Language Models
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 750792ca-cc5a-478a-8f98-0e2f4d552475 · outbound
Enhancing LLMs' Reasoning-Intensive Multimedia Search Capabilities through Fine-Tuning and Reinforcement Learning Unresolved cited work
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0d0d412e-0076-4204-94fe-b559e03ca013 · outbound
Enhancing LLMs' Reasoning-Intensive Multimedia Search Capabilities through Fine-Tuning and Reinforcement Learning CFBenchmark: Chinese Financial Assistant Benchmark for Large Language Model
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 36cf3bfc-5601-4f4d-bc80-a95622c814d3 · outbound
Enhancing LLMs' Reasoning-Intensive Multimedia Search Capabilities through Fine-Tuning and Reinforcement Learning An Agent Framework for Real-Time Financial Information Searching with Large Language Models
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d0bd0eeb-95f8-4056-8104-7f10d90f1f7a · outbound
Enhancing LLMs' Reasoning-Intensive Multimedia Search Capabilities through Fine-Tuning and Reinforcement Learning BLIP-2: Bootstrapping Language-Image Pre-training with Frozen Image Encoders and Large Language Models
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f08b0495-f438-4f42-bebb-920bf9c62329 · outbound
Enhancing LLMs' Reasoning-Intensive Multimedia Search Capabilities through Fine-Tuning and Reinforcement Learning Guerreiro, Ricardo Rei, and André F
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 73b135fb-f751-41b0-9faa-c6c38078de3a · outbound
Enhancing LLMs' Reasoning-Intensive Multimedia Search Capabilities through Fine-Tuning and Reinforcement Learning Scalable Diffusion Models with Transformers
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c09c6818-50c4-4fb2-984c-f7c2a686d488 · outbound
Enhancing LLMs' Reasoning-Intensive Multimedia Search Capabilities through Fine-Tuning and Reinforcement Learning Comparing Traditional and LLM-based Search for Consumer Choice: A Randomized Experiment
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7cea74e3-a911-4837-9650-50ffeb60436d · outbound
Enhancing LLMs' Reasoning-Intensive Multimedia Search Capabilities through Fine-Tuning and Reinforcement Learning ChatDev: Communicative Agents for Software Development
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation edd306ae-02fe-48cf-aa94-51706cb421df · outbound
Enhancing LLMs' Reasoning-Intensive Multimedia Search Capabilities through Fine-Tuning and Reinforcement Learning Gemini: A Family of Highly Capable Multimodal Models
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b6349f8a-556b-4ac1-9675-7700c1b0c9c4 · outbound
Enhancing LLMs' Reasoning-Intensive Multimedia Search Capabilities through Fine-Tuning and Reinforcement Learning Qwen2.5-32B: Leveraging Self-Consistent Tool-Integrated Reasoning for Bengali Mathematical Olympiad Problem Solving
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 75960073-c17a-4b22-a1eb-b9bd92624d67 · outbound
Enhancing LLMs' Reasoning-Intensive Multimedia Search Capabilities through Fine-Tuning and Reinforcement Learning AgentRM: Enhancing Agent Generalization with Reward Modeling
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f6c886b5-483f-4cf4-8968-c21974c3a241 · outbound
Enhancing LLMs' Reasoning-Intensive Multimedia Search Capabilities through Fine-Tuning and Reinforcement Learning LLaMA: Open and Efficient Foundation Language Models
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3425af78-980c-4011-bb1d-316548ee1f9e · outbound
Enhancing LLMs' Reasoning-Intensive Multimedia Search Capabilities through Fine-Tuning and Reinforcement Learning Qwen2 Technical Report
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e473578f-6892-4132-860e-fda9dfc47eca · outbound
Enhancing LLMs' Reasoning-Intensive Multimedia Search Capabilities through Fine-Tuning and Reinforcement Learning When Search Engine Services meet Large Language Models: Visions and Challenges
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3a082e5b-f3c2-4673-82a7-e5dce44208d8 · outbound
Enhancing LLMs' Reasoning-Intensive Multimedia Search Capabilities through Fine-Tuning and Reinforcement Learning KoLA: Carefully Benchmarking World Knowledge of Large Language Models
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 96fa39a9-a874-41d9-b13b-7de76a7f2b20 · outbound
Enhancing LLMs' Reasoning-Intensive Multimedia Search Capabilities through Fine-Tuning and Reinforcement Learning Rethinking Prompt-based Debiasing in Large Language Models
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 295697fa-3ccd-42ca-879d-7ef5f9d16560 · outbound
Enhancing LLMs' Reasoning-Intensive Multimedia Search Capabilities through Fine-Tuning and Reinforcement Learning Supervised Fine-Tuning Achieve Rapid Task Adaption Via Alternating Attention Head Activation Patterns
Reference 36
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cc5701c9-e022-4f44-91c4-62ebc5f5da33 · outbound
Enhancing LLMs' Reasoning-Intensive Multimedia Search Capabilities through Fine-Tuning and Reinforcement Learning FinLLMs: A Framework for Financial Reasoning Dataset Generation with Large Language Models
Reference 37
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4a305ed3-4c3d-454d-af58-5eb73bb3d940 · outbound
Enhancing LLMs' Reasoning-Intensive Multimedia Search Capabilities through Fine-Tuning and Reinforcement Learning LlamaFactory: Unified Efficient Fine-Tuning of 100+ Language Models
Reference 39
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a46bd610-488d-4298-9323-952222564da3 · outbound
Enhancing LLMs' Reasoning-Intensive Multimedia Search Capabilities through Fine-Tuning and Reinforcement Learning Recurrence-Enhanced Vision-and-Language Transformers for Robust Multimodal Document Retrieval
Reference 2025
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation bf6c5fa1-85b2-49ab-a248-e5411c7e1338 · inbound
Reinforcement Fine-Tuning for Reasoning towards Multi-Step Multi-Source Search in Large Language Models Enhancing LLMs' Reasoning-Intensive Multimedia Search Capabilities through Fine-Tuning and Reinforcement Learning
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.