Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-12T19:52:34.445566Z
Paper Citation Record · LEDGER
As of 12 August 2026, this Paper Citation Record lists 32 of 32 outbound references and 0 inbound Pith citation observations for arXiv:2411.10252.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-12T19:52:34.445566Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-12T06:34:41.77262+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links
A source-named dated measurement, never combined with another source.
Source: cited_works
32 of 32 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 6bab4c03-8d17-4f37-9210-c0c87e688353 · outbound
Visual-Linguistic Agent: Towards Collaborative Contextual Object Reasoning Vqa: Visual question answering
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 8164ff06-df70-402e-a88e-394ba00bd617 · outbound
Visual-Linguistic Agent: Towards Collaborative Contextual Object Reasoning End-to- end object detection with transformers
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 497d0734-49e0-43e4-98b1-ea480eb2aa38 · outbound
Visual-Linguistic Agent: Towards Collaborative Contextual Object Reasoning Spatial memory for context reasoning in object detection
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 4c7d22df-37b1-4823-9f10-fb6af8f61c46 · outbound
Visual-Linguistic Agent: Towards Collaborative Contextual Object Reasoning PaLM-E: An Embodied Multimodal Language Model
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 57f6ee1e-1a6e-45f4-8e5b-ed80f5d3c9e8 · outbound
Visual-Linguistic Agent: Towards Collaborative Contextual Object Reasoning Relation networks for object detection
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 7d996949-d5e4-4ee2-96a7-865a816aec46 · outbound
Visual-Linguistic Agent: Towards Collaborative Contextual Object Reasoning Dac-detr: Divide the attention layers and conquer
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 2a9169a6-ae50-4225-b18f-51e92addc00a · outbound
Visual-Linguistic Agent: Towards Collaborative Contextual Object Reasoning Hugging face
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 748f5a05-3d9d-431a-81fc-ac3f7c2dff51 · outbound
Visual-Linguistic Agent: Towards Collaborative Contextual Object Reasoning Capabilities of Large Language Models in Control Engineering: A Benchmark Study on GPT-4, Claude 3 Opus, and Gemini 1.0 Ultra
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 16e860df-6a81-4d28-aa0d-6d050ea154ce · outbound
Visual-Linguistic Agent: Towards Collaborative Contextual Object Reasoning YOLOv11: An Overview of the Key Architectural Enhancements
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1c198296-dd7d-45cd-88c7-d7975601b408 · outbound
Visual-Linguistic Agent: Towards Collaborative Contextual Object Reasoning Seed-bench: Bench- marking multimodal large language models
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 22fb73db-7ecb-4055-82cf-01c456cb688d · outbound
Visual-Linguistic Agent: Towards Collaborative Contextual Object Reasoning Blip-2: Bootstrapping language-image pre-training with frozen image encoders and large language models
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5456ba81-9a04-4517-8dda-7cf406361a37 · outbound
Visual-Linguistic Agent: Towards Collaborative Contextual Object Reasoning Microsoft coco: Common objects in context
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fd188497-08a5-47c4-92bc-71550d0c102d · outbound
Visual-Linguistic Agent: Towards Collaborative Contextual Object Reasoning Visual Instruction Tuning
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation eba186b5-0727-42b6-9274-7cccfc43bb54 · outbound
Visual-Linguistic Agent: Towards Collaborative Contextual Object Reasoning Cigar: Cross-modality graph reasoning for domain adaptive object detection
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation bc87118b-775a-4597-aa73-379630865c19 · outbound
Visual-Linguistic Agent: Towards Collaborative Contextual Object Reasoning Rt-gcn: Gaussian-based spatiotemporal graph convolutional network for robust traffic prediction.In- formation Fusion, 102:102078, 2024
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 0e8b8201-ca0b-4297-a91e-492235df3795 · outbound
Visual-Linguistic Agent: Towards Collaborative Contextual Object Reasoning Compositional chain-of-thought prompting for large multimodal models
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation df02a91d-858c-4a10-898f-251054fe041b · outbound
Visual-Linguistic Agent: Towards Collaborative Contextual Object Reasoning Improving multimodal datasets with image captioning
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1c40f5f1-9591-4320-aa40-0ad5103418ea · outbound
Visual-Linguistic Agent: Towards Collaborative Contextual Object Reasoning Faster r-cnn: Towards real-time object detection with region proposal networks
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 9a08821b-b50d-4e3e-9799-65fc3ddab5ab · outbound
Visual-Linguistic Agent: Towards Collaborative Contextual Object Reasoning Hugginggpt: Solving ai tasks with chatgpt and its friends in hugging face
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c65c301f-da18-4661-95fa-6e4a9b7e33b4 · outbound
Visual-Linguistic Agent: Towards Collaborative Contextual Object Reasoning Vipergpt: Visual inference via python execution for reasoning
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8d59b5e0-b21d-4934-aa4d-f180653602f8 · outbound
Visual-Linguistic Agent: Towards Collaborative Contextual Object Reasoning LLaMA: Open and Efficient Foundation Language Models
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation beb84e66-a76a-482a-ab85-60f7be330d96 · outbound
Visual-Linguistic Agent: Towards Collaborative Contextual Object Reasoning Sw-yolox: A yolox-based real-time pedestrian detector with shift window- mixed attention mechanism
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation f56eda0a-6fef-4042-ae26-7cbf19d71c6b · outbound
Visual-Linguistic Agent: Towards Collaborative Contextual Object Reasoning Robust motor- cycle helmet detection in real-world scenarios: Using co- detr and minority class enhancement
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation c256903f-50eb-447b-8689-ebdac6f56ea5 · outbound
Visual-Linguistic Agent: Towards Collaborative Contextual Object Reasoning Chain-of-thought prompting elicits reasoning in large lan- guage models
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1f29ecd7-33d0-4a94-81a3-511b922490d9 · outbound
Visual-Linguistic Agent: Towards Collaborative Contextual Object Reasoning Spatial-aware graph relation network for large-scale object detection
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation ba35c89b-9d60-4587-ac8f-b00bf13a94d0 · outbound
Visual-Linguistic Agent: Towards Collaborative Contextual Object Reasoning Dino: Detr with improved denoising anchor boxes for end-to-end object de- tection
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 862dfc8f-5428-4ee9-9e01-39481688e8ba · outbound
Visual-Linguistic Agent: Towards Collaborative Contextual Object Reasoning DINO: DETR with Improved DeNoising Anchor Boxes for End-to-End Object Detection
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 13e55193-fec1-457e-ab7d-3578acd0cec5 · outbound
Visual-Linguistic Agent: Towards Collaborative Contextual Object Reasoning Ms-detr: Efficient detr training with mixed supervision
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation a7856eac-afc3-44f8-9bc2-2cb943cffae9 · outbound
Visual-Linguistic Agent: Towards Collaborative Contextual Object Reasoning Rgrn: Relation-aware graph reasoning network for object de- tection
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 4fdae871-b264-4c8c-b8bb-3cd709c50eb5 · outbound
Visual-Linguistic Agent: Towards Collaborative Contextual Object Reasoning Semantic relation reasoning for shot- stable few-shot object detection
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 11a22183-2bb6-4e6e-a4bb-d76a42d65ae9 · outbound
Visual-Linguistic Agent: Towards Collaborative Contextual Object Reasoning MiniGPT-4: Enhancing Vision-Language Understanding with Advanced Large Language Models
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6f9c0d52-e485-4146-8a64-2404eadc8adc · outbound
Visual-Linguistic Agent: Towards Collaborative Contextual Object Reasoning An efficient two-state gru based on feature attention mechanism for sen- timent analysis
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
No inbound Pith citation observations are available.