Pith. sign in

Paper Citation Record · LEDGER

From Answers to Rationales: Self-Aligning Multimodal Reasoning with Answer-Oriented Chain-of-Thought

As of 7 August 2026, this Paper Citation Record lists 98 of 98 outbound references and 0 inbound Pith citation observations for arXiv:2507.02984.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2507.02984 v2

Coverage vector

measured 98 of 98 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-06T21:19:42.977063Z

measured 98 of 98 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

98 of 98 outbound references displayed

  • verified exact0
  • verified fuzzy22
  • unresolved75
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch1

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation f416be21-c02a-4b9b-af2c-fcf74533376d · outbound

This paper cites MiniGPT-v2: large language model as a unified interface for vision-language multi-task learning.

From Answers to Rationales: Self-Aligning Multimodal Reasoning with Answer-Oriented Chain-of-Thought MiniGPT-v2: large language model as a unified interface for vision-language multi-task learning

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-06T21:19:33.852245Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:19:33.852245Z digest=sha256:5d2e2696313e07e038cf5f34f812b91868712dfc3c5bfa8c01cefebcb584e3de

Observation 6b577f96-c747-4113-88ce-7e575130aeed · outbound

This paper cites Shikra: Unleashing Multimodal LLM's Referential Dialogue Magic.

From Answers to Rationales: Self-Aligning Multimodal Reasoning with Answer-Oriented Chain-of-Thought Shikra: Unleashing Multimodal LLM's Referential Dialogue Magic

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-06T21:19:33.912791Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:19:33.912791Z digest=sha256:4552915379b7ab32e61854b24a25cbac31d8cf4e681bc7765946822cb6b85834

Observation 13ec6268-f581-4eb7-a91e-dd40699eb75b · outbound

This paper cites Kosmos-2: Grounding Multimodal Large Language Models to the World.

From Answers to Rationales: Self-Aligning Multimodal Reasoning with Answer-Oriented Chain-of-Thought Kosmos-2: Grounding Multimodal Large Language Models to the World

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-06T21:19:34.013688Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:19:34.013688Z digest=sha256:338a44a3848f1a4a8278ef6ae1e75e7880643771355549d4a2123263d43b6c97

Observation fe18d58a-759a-4aeb-a4ab-e676cca4ece0 · outbound

This paper cites CogVLM2: Visual Language Models for Image and Video Understanding.

From Answers to Rationales: Self-Aligning Multimodal Reasoning with Answer-Oriented Chain-of-Thought CogVLM2: Visual Language Models for Image and Video Understanding

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-06T21:19:34.077710Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:19:34.077710Z digest=sha256:02c480c4ed507a0537e3f6f80ca91891139d4bcdac43fa8d0aeb29675adecf95

Observation 9f277b7f-a5e5-4af2-999b-fbd7e98ae66f · outbound

This paper cites In: International Conference on Machine Learning, pp.

From Answers to Rationales: Self-Aligning Multimodal Reasoning with Answer-Oriented Chain-of-Thought In: International Conference on Machine Learning, pp

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-06T21:19:34.177966Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:19:34.177966Z digest=sha256:5ad7d5efda5fcbe553b276c4bb378b7ac7901df9fa25a4f2821d05644cd800e3

Observation 1f93425b-bfc9-47e1-a7d1-4d58c7ce2b8d · outbound

This paper cites In: Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition, pp.

From Answers to Rationales: Self-Aligning Multimodal Reasoning with Answer-Oriented Chain-of-Thought In: Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition, pp

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-06T21:19:34.243662Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:19:34.243662Z digest=sha256:9e5bcc689c8c764acbc68df9f509efdd2d7f846bff2da62c4f97fd6862901c26

Observation 9a18fa4d-9cd2-47c2-b3ad-e2f59471e96d · outbound

This paper cites In: Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition, pp.

From Answers to Rationales: Self-Aligning Multimodal Reasoning with Answer-Oriented Chain-of-Thought In: Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition, pp

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-06T21:19:34.353123Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:19:34.353123Z digest=sha256:1ca604e1918c5785af34f3e151034ec0d7be6c9a87e27ca40668a1feb03ce5ff

Observation 927fd270-58f6-46c2-ade2-ded5b8f8ed54 · outbound

This paper cites SPHINX: The Joint Mixing of Weights, Tasks, and Visual Embeddings for Multi-modal Large Language Models.

From Answers to Rationales: Self-Aligning Multimodal Reasoning with Answer-Oriented Chain-of-Thought SPHINX: The Joint Mixing of Weights, Tasks, and Visual Embeddings for Multi-modal Large Language Models

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-06T21:19:34.422630Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:19:34.422630Z digest=sha256:31fe3341d23d7fc545192a2115951d0e4a938fe4199e3934bcbbef82c079ecfa

Observation 3d524dfb-0e36-4805-ba3e-b1bbf50ed920 · outbound

This paper cites How Far Are We to GPT-4V? Closing the Gap to Commercial Multimodal Models with Open-Source Suites.

From Answers to Rationales: Self-Aligning Multimodal Reasoning with Answer-Oriented Chain-of-Thought How Far Are We to GPT-4V? Closing the Gap to Commercial Multimodal Models with Open-Source Suites

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-06T21:19:34.505690Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:19:34.505690Z digest=sha256:a2faea46dcfaf808bbd297b65c8c259d193f9265afe44b796406f84340123ba8

Observation 6c9756ba-a469-4149-b741-3c25008706bd · outbound

This paper cites DeepSeek-VL: Towards Real-World Vision-Language Understanding.

From Answers to Rationales: Self-Aligning Multimodal Reasoning with Answer-Oriented Chain-of-Thought DeepSeek-VL: Towards Real-World Vision-Language Understanding

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-06T21:19:34.629464Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:19:34.629464Z digest=sha256:98c795b4a0ebe249aa17ddbe31a57f28142aec60c9d448fabfc46864ecdbd4a9

Observation e3ec2494-dc05-4060-ae76-a6cb5f1f5a56 · outbound

This paper cites Yi: Open Foundation Models by 01.AI.

From Answers to Rationales: Self-Aligning Multimodal Reasoning with Answer-Oriented Chain-of-Thought Yi: Open Foundation Models by 01.AI

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-06T21:19:34.703051Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:19:34.703051Z digest=sha256:f4ac00e27bfc02da41a490c878c51ed9d114c991b444e98e6a687f5034255f24

Observation cc9022a1-7334-4aff-8f1d-9af90499af99 · outbound

This paper cites https://llava-vl.github.

From Answers to Rationales: Self-Aligning Multimodal Reasoning with Answer-Oriented Chain-of-Thought https://llava-vl.github

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-06T21:19:34.772600Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:19:34.772600Z digest=sha256:84ad7f310b707041e1d108de79d0ff2d82e93b10010060253fe4050d926b990a

Observation 055f57a2-4f6f-4f21-b0e6-8ab138540039 · outbound

This paper cites MiniCPM-V: A GPT-4V Level MLLM on Your Phone.

From Answers to Rationales: Self-Aligning Multimodal Reasoning with Answer-Oriented Chain-of-Thought MiniCPM-V: A GPT-4V Level MLLM on Your Phone

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-06T21:19:34.859505Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:19:34.859505Z digest=sha256:1a5a9eb683bbf12d4758da6f3722874617cecac5006cf257222cf6d28641278c

Observation 2422e848-f067-473f-bb2f-4a07fd237b05 · outbound

This paper cites Qwen2-VL: Enhancing Vision-Language Model's Perception of the World at Any Resolution.

From Answers to Rationales: Self-Aligning Multimodal Reasoning with Answer-Oriented Chain-of-Thought Qwen2-VL: Enhancing Vision-Language Model's Perception of the World at Any Resolution

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-06T21:19:34.937712Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:19:34.937712Z digest=sha256:06928fa8154435dd1d2ae3eef976e298aa112a8479aad437473b20587c72593e

Observation 4708ab48-5e6b-485b-9155-412f188f8d23 · outbound

This paper cites AMBER: An LLM-free Multi-dimensional Benchmark for MLLMs Hallucination Evaluation.

From Answers to Rationales: Self-Aligning Multimodal Reasoning with Answer-Oriented Chain-of-Thought AMBER: An LLM-free Multi-dimensional Benchmark for MLLMs Hallucination Evaluation

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-06T21:19:35.022823Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:19:35.022823Z digest=sha256:72d570889f7e3df77ef1bbdde4043af7da2a81d7dca0590747f5f93485bf6b30

Observation 7900ed98-6e53-4bac-8741-2e33725432fe · outbound

This paper cites NeurIPS (2023).

From Answers to Rationales: Self-Aligning Multimodal Reasoning with Answer-Oriented Chain-of-Thought NeurIPS (2023)

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-06T21:19:35.110223Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:19:35.110223Z digest=sha256:e3f8c13489d3eeabb321ce10d1eb70dfdaa800ca40fd7864b4d2985f5179da8d

Observation b2ad41e5-2682-428e-b7ab-be0596cb1577 · outbound

This paper cites MME: A Comprehensive Evaluation Benchmark for Multimodal Large Language Models.

From Answers to Rationales: Self-Aligning Multimodal Reasoning with Answer-Oriented Chain-of-Thought MME: A Comprehensive Evaluation Benchmark for Multimodal Large Language Models

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-06T21:19:35.203253Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:19:35.203253Z digest=sha256:c12e232fad9f78f9dd5ab60e5924bd9a9633b1156620fde2bd29f0d08cfb40cd

Observation e4225a55-a0a2-487c-9e2e-1be11a7eea43 · outbound

This paper cites In: Proceedings of the IEEE International Conference on Computer Vision, pp.

From Answers to Rationales: Self-Aligning Multimodal Reasoning with Answer-Oriented Chain-of-Thought In: Proceedings of the IEEE International Conference on Computer Vision, pp

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-06T21:19:35.260301Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:19:35.260301Z digest=sha256:24daceca95f5aeafc60cecbfbdcdacc13a7e8feaf2a277719eca5ccffb9c5323

Observation 26b4a820-0ff1-4e90-b6c6-3dfd42f199ec · outbound

This paper cites In: Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition, pp.

From Answers to Rationales: Self-Aligning Multimodal Reasoning with Answer-Oriented Chain-of-Thought In: Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition, pp

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-06T21:19:35.355158Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:19:35.355158Z digest=sha256:431e85293a70e6858b5a72a8ca12d50aa820ae58bf657d75b233cb6ca6ebefd2

Observation 04819b31-00e6-4373-a1cb-084c8649815d · outbound

This paper cites MM-Vet: Evaluating Large Multimodal Models for Integrated Capabilities.

From Answers to Rationales: Self-Aligning Multimodal Reasoning with Answer-Oriented Chain-of-Thought MM-Vet: Evaluating Large Multimodal Models for Integrated Capabilities

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-06T21:19:35.419588Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:19:35.419588Z digest=sha256:3820b5b32cd0ed86d4f25dcd7ee71ca4e88e1e2af649360f756835e8da4c64b1

Observation 89997c7a-e103-4a9d-ad83-bba8518244d1 · outbound

This paper cites In: International Conference on Learning Representations (ICLR) (2024).

From Answers to Rationales: Self-Aligning Multimodal Reasoning with Answer-Oriented Chain-of-Thought In: International Conference on Learning Representations (ICLR) (2024)

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-06T21:19:35.492421Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:19:35.492421Z digest=sha256:fe3901738cbd45d8ef60cb606a2c249f914fa6d9b39c25e3c06f93a7029eeb18

Observation 284e0a8b-399e-456d-b674-0fcf5fb39f21 · outbound

This paper cites ACL (2024).

From Answers to Rationales: Self-Aligning Multimodal Reasoning with Answer-Oriented Chain-of-Thought ACL (2024)

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-06T21:19:35.556816Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:19:35.556816Z digest=sha256:92da9bf4cbd65f9d2be38e0ac9de6d058f9662f58095166d82da22ae21ccef39

Observation f07419c3-9f53-422c-82e6-d0314e8f14f2 · outbound

This paper cites NeurIPS (2022).

From Answers to Rationales: Self-Aligning Multimodal Reasoning with Answer-Oriented Chain-of-Thought NeurIPS (2022)

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-06T21:19:35.650695Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:19:35.650695Z digest=sha256:e094d91cc5c59a617b671655fc5749f9df1a41ba5b05f9c30d455f67256c8805

Observation d64192ec-5a48-479b-98c3-4c1ae57b6b90 · outbound

This paper cites Multimodal Chain-of-Thought Reasoning in Language Models.

From Answers to Rationales: Self-Aligning Multimodal Reasoning with Answer-Oriented Chain-of-Thought Multimodal Chain-of-Thought Reasoning in Language Models

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-06T21:19:35.695869Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:19:35.695869Z digest=sha256:0b1def0d117931fee39a4fc03c5672c419ec52c009a1aad7980a30315eb51155

Observation 5dd3c81e-b9b3-49bf-8400-c58b711e5f56 · outbound

This paper cites NeurIPS (2023).

From Answers to Rationales: Self-Aligning Multimodal Reasoning with Answer-Oriented Chain-of-Thought NeurIPS (2023)

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-06T21:19:35.779999Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:19:35.779999Z digest=sha256:73b270b2d903d1f2cdb05e40181461acb7632172b9270b1979a5178af10e4aa4

Observation 089aa579-19a6-41ed-875b-9525f7e724b3 · outbound

This paper cites In: Proceedings of the AAAI Conference on Artificial Intelligence, vol.

From Answers to Rationales: Self-Aligning Multimodal Reasoning with Answer-Oriented Chain-of-Thought In: Proceedings of the AAAI Conference on Artificial Intelligence, vol

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-06T21:19:35.846251Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:19:35.846251Z digest=sha256:2604e18df941331823743061656debf09d07f6001fa7896651e806cc15cd3d4c

Observation 9fe9cbd2-4458-4d38-98ff-8964cf14f0fc · outbound

This paper cites ACM MM (2024) 16.

From Answers to Rationales: Self-Aligning Multimodal Reasoning with Answer-Oriented Chain-of-Thought ACM MM (2024) 16

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-06T21:19:35.912123Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:19:35.912123Z digest=sha256:a8570596a17bb84585807b6dfe99e74f34c2377e0a934f2106ceb8e18f9d6994

Observation 207edfa8-d3a2-4a54-a274-d4ba0697749a · outbound

This paper cites In: CVPR (2024).

From Answers to Rationales: Self-Aligning Multimodal Reasoning with Answer-Oriented Chain-of-Thought In: CVPR (2024)

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-06T21:19:35.992476Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:19:35.992476Z digest=sha256:64fb70926d9bac630001e85ae938c1f0f2a50a9cb7c2a094d79cf14f71eb0cf0

Observation 518bea90-bf82-478e-9b14-1bafb206c194 · outbound

This paper cites The Role of Chain-of-Thought in Complex Vision-Language Reasoning Task.

From Answers to Rationales: Self-Aligning Multimodal Reasoning with Answer-Oriented Chain-of-Thought The Role of Chain-of-Thought in Complex Vision-Language Reasoning Task

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-06T21:19:36.092686Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:19:36.092686Z digest=sha256:5aca3f0f094c4ac04f84a59b847a8191865bb1a8a6feac18856e18aceec59b02

Observation 5b546750-8f74-4020-b4b0-484e32544102 · outbound

This paper cites Enhancing Large Vision Language Models with Self-Training on Image Comprehension.

From Answers to Rationales: Self-Aligning Multimodal Reasoning with Answer-Oriented Chain-of-Thought Enhancing Large Vision Language Models with Self-Training on Image Comprehension

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-06T21:19:36.169281Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:19:36.169281Z digest=sha256:e65857e7568a80297c0cd74b6f4b6fdf89a4931da726094bd67d63c7631bba26

Observation dcc6de35-d0ab-40e1-9116-b0127168f179 · outbound

This paper cites ISR-DPO: Aligning Large Multimodal Models for Videos by Iterative Self-Retrospective DPO.

From Answers to Rationales: Self-Aligning Multimodal Reasoning with Answer-Oriented Chain-of-Thought ISR-DPO: Aligning Large Multimodal Models for Videos by Iterative Self-Retrospective DPO

Reference 31

Resolution
metadata mismatch
local_arxiv, observed 2026-08-06T21:19:43.391954Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T21:19:36.242869Z digest=sha256:e3883d17d2abe7b4d64894cd0a14b32315af799e53601baa4ccb526c5e1dae54

Observation 898c1883-d398-4829-a80a-277d07070c79 · outbound

This paper cites Enhancing Visual-Language Modality Alignment in Large Vision Language Models via Self-Improvement.

From Answers to Rationales: Self-Aligning Multimodal Reasoning with Answer-Oriented Chain-of-Thought Enhancing Visual-Language Modality Alignment in Large Vision Language Models via Self-Improvement

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-06T21:19:36.307866Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:19:36.307866Z digest=sha256:738e3b5133d78b14fce4af90a7352a3839c01f12f9fb968fb93c832cc926c790

Observation f00da12f-f109-47a1-9460-1bdec9af2d54 · outbound

This paper cites Beyond Human Data: Aligning Multimodal Large Language Models by Iterative Self-Evolution.

From Answers to Rationales: Self-Aligning Multimodal Reasoning with Answer-Oriented Chain-of-Thought Beyond Human Data: Aligning Multimodal Large Language Models by Iterative Self-Evolution

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-06T21:19:36.373666Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:19:36.373666Z digest=sha256:4d6ed9954161509ee01989334ad776ca2dcb8f42c670c2cfe1bd2fc2d8a630fa

Observation c9fb9dea-dadb-44bc-9070-4029abaf468e · outbound

This paper cites In: Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition, pp.

From Answers to Rationales: Self-Aligning Multimodal Reasoning with Answer-Oriented Chain-of-Thought In: Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition, pp

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-06T21:19:36.451178Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:19:36.451178Z digest=sha256:b77e71a2e034466a252d4a99ea98a2a5e2cfd2ec7917febe81e9750c9fc01716

Observation 56362f2f-5a59-4b3c-83fc-daa77b6a2a8b · outbound

This paper cites In: Computer Vision–ECCV 2016: 14th European Conference, Amsterdam, The Netherlands, October 11–14, 2016, Proceedings, Part IV 14, pp.

From Answers to Rationales: Self-Aligning Multimodal Reasoning with Answer-Oriented Chain-of-Thought In: Computer Vision–ECCV 2016: 14th European Conference, Amsterdam, The Netherlands, October 11–14, 2016, Proceedings, Part IV 14, pp

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-06T21:19:36.528716Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:19:36.528716Z digest=sha256:759155529eb682d672d547bdf832fd54d7e9350ce82f579a4fc44211146391d2

Observation 18401eb3-51b1-4d45-bc32-17a14552b1ca · outbound

This paper cites an unresolved cited work.

From Answers to Rationales: Self-Aligning Multimodal Reasoning with Answer-Oriented Chain-of-Thought Unresolved cited work

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-06T21:19:36.608901Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:19:36.608901Z digest=sha256:c6fa0e6c12de68e4a64a9a43e6c8f68377b49ffc62c01f04058c769f84d8945a

Observation e0aec378-e836-4353-a392-163c639c9984 · outbound

This paper cites Measuring Mathematical Problem Solving With the MATH Dataset.

From Answers to Rationales: Self-Aligning Multimodal Reasoning with Answer-Oriented Chain-of-Thought Measuring Mathematical Problem Solving With the MATH Dataset

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-06T21:19:36.653815Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:19:36.653815Z digest=sha256:79bca47e0cb853d5c49bcb97f18d21fb03605c8df723277bdbf16b1340e8db53

Observation b1d80a42-65fc-418e-a042-88622768a4dd · outbound

This paper cites In: European Conference on Computer Vision, pp.

From Answers to Rationales: Self-Aligning Multimodal Reasoning with Answer-Oriented Chain-of-Thought In: European Conference on Computer Vision, pp

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-06T21:19:36.720980Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:19:36.720980Z digest=sha256:6549e75831d1648bc85294cc3a94fb6b77dc62375155011220845cd2de0a17e4

Observation 90347a6c-eb5a-4ac5-9b91-3134c0d0cba6 · outbound

This paper cites CoCoT: Contrastive Chain-of-Thought Prompting for Large Multimodal Models with Multiple Image Inputs.

From Answers to Rationales: Self-Aligning Multimodal Reasoning with Answer-Oriented Chain-of-Thought CoCoT: Contrastive Chain-of-Thought Prompting for Large Multimodal Models with Multiple Image Inputs

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-06T21:19:36.765868Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:19:36.765868Z digest=sha256:bcbc3d5fdad42f1e3c072958f3fce60423223e701c01f41b43496dbb9025f08b

Observation ce0619b4-d8bf-4350-9aa7-441eacbc1a92 · outbound

This paper cites Visual CoT: Advancing Multi-Modal Language Models with a Comprehensive Dataset and Benchmark for Chain-of-Thought Reasoning.

From Answers to Rationales: Self-Aligning Multimodal Reasoning with Answer-Oriented Chain-of-Thought Visual CoT: Advancing Multi-Modal Language Models with a Comprehensive Dataset and Benchmark for Chain-of-Thought Reasoning

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-06T21:19:36.818352Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:19:36.818352Z digest=sha256:9f7ba79051824fa0c04ccb1c86aaba9482d7839325eb61aa4bc7c348d8defb0e

Observation 488629da-b6f5-4919-b848-b543690d8c37 · outbound

This paper cites Journal of machine learning research 21(140), 1–67 (2020).

From Answers to Rationales: Self-Aligning Multimodal Reasoning with Answer-Oriented Chain-of-Thought Journal of machine learning research 21(140), 1–67 (2020)

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:19:45.279871Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T21:19:36.869481Z digest=sha256:20987abd0943be1719d9f99fa19fd34b0a3ddfdd092b076fda36943f8469fac2

Observation a9ac94fc-c90b-4f57-b292-963270d15c1e · outbound

This paper cites : Training language models to follow instructions with human feedback.

From Answers to Rationales: Self-Aligning Multimodal Reasoning with Answer-Oriented Chain-of-Thought : Training language models to follow instructions with human feedback

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:19:45.268677Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T21:19:36.945519Z digest=sha256:959af2b103bf55a9bfa347da06676302eb4c1244c22c99f22264929fc5b0ae0e

Observation 8ad19dd6-f942-4ace-90d9-dab0c91f6bc5 · outbound

This paper cites LLaVA-NeXT-Interleave: Tackling Multi-image, Video, and 3D in Large Multimodal Models.

From Answers to Rationales: Self-Aligning Multimodal Reasoning with Answer-Oriented Chain-of-Thought LLaVA-NeXT-Interleave: Tackling Multi-image, Video, and 3D in Large Multimodal Models

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-06T21:19:36.992645Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:19:36.992645Z digest=sha256:8d9f9e2a4c6be3456349b96b334f7d59e7a8ba4fed3710e7b87cbb1c1aa498fd

Observation 009094eb-d556-4465-a87d-d4e5bc3d9846 · outbound

This paper cites In: Proceedings of the 37th International Conference on Neural Information Processing Systems (2024).

From Answers to Rationales: Self-Aligning Multimodal Reasoning with Answer-Oriented Chain-of-Thought In: Proceedings of the 37th International Conference on Neural Information Processing Systems (2024)

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:19:45.257314Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T21:19:37.046109Z digest=sha256:a230ef43205aaa31b46a04a0a309c11d2b7951dbf868c5d68519800505aaf200

Observation a56f1546-e8ca-4dee-91e8-64d5b58d6023 · outbound

This paper cites https://openai.com/research/ gpt-4v-system-card.

From Answers to Rationales: Self-Aligning Multimodal Reasoning with Answer-Oriented Chain-of-Thought https://openai.com/research/ gpt-4v-system-card

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:19:45.245815Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T21:19:37.124103Z digest=sha256:3542476361714b02ece9a0dec3f9d184d3d91785e8b99dcf39fff88ba0e3ddd7

Observation 1f54acb1-611e-49c7-ba95-7ebfe0a7f823 · outbound

This paper cites Direct Language Model Alignment from Online AI Feedback.

From Answers to Rationales: Self-Aligning Multimodal Reasoning with Answer-Oriented Chain-of-Thought Direct Language Model Alignment from Online AI Feedback

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-06T21:19:37.182865Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:19:37.182865Z digest=sha256:f9f6e434ff7e74eea5f57bb0506b11b26d6854a6773cfd9c004443aceaed306e

Observation 7c78759c-5123-4548-b72a-5f8e56dc45c4 · outbound

This paper cites Self-Rewarding Language Models.

From Answers to Rationales: Self-Aligning Multimodal Reasoning with Answer-Oriented Chain-of-Thought Self-Rewarding Language Models

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-06T21:19:37.231697Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:19:37.231697Z digest=sha256:e4b2cc3dd1452d1e2f4e87b0a7f8a40e5c77ed16ef4e62933dc52d0322bfba91

Observation 15b86385-82d0-4368-b077-1c142b2587e1 · outbound

This paper cites Human Alignment of Large Language Models through Online Preference Optimisation.

From Answers to Rationales: Self-Aligning Multimodal Reasoning with Answer-Oriented Chain-of-Thought Human Alignment of Large Language Models through Online Preference Optimisation

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-06T21:19:37.284453Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:19:37.284453Z digest=sha256:5a37e06ab9aeab7921d082edbbba34bda15d87537935a0fb06fda9c94cbc03ee

Observation 67e203a4-c003-42bd-a4d7-b46b6906879b · outbound

This paper cites RLHF Workflow: From Reward Modeling to Online RLHF.

From Answers to Rationales: Self-Aligning Multimodal Reasoning with Answer-Oriented Chain-of-Thought RLHF Workflow: From Reward Modeling to Online RLHF

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-06T21:19:37.331130Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:19:37.331130Z digest=sha256:8fe640393fb9fe091276031de7258c670de58186bf6d5ac6b77b6cc4699aae2f

Observation 884b4a97-d57d-43a0-98d2-87d6e9ac026f · outbound

This paper cites Self-Play Fine-Tuning Converts Weak Language Models to Strong Language Models.

From Answers to Rationales: Self-Aligning Multimodal Reasoning with Answer-Oriented Chain-of-Thought Self-Play Fine-Tuning Converts Weak Language Models to Strong Language Models

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-06T21:19:37.398160Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:19:37.398160Z digest=sha256:7941ed87a9d2d6f25a02dbe7d2d868b07d97122f09a02a4891c847a66bf5e0e0

Observation 07f27bc1-2ecd-417a-b148-5556b19f72db · outbound

This paper cites Advances in Neural Information Processing Systems 35, 15476–15488 (2022).

From Answers to Rationales: Self-Aligning Multimodal Reasoning with Answer-Oriented Chain-of-Thought Advances in Neural Information Processing Systems 35, 15476–15488 (2022)

Reference 51

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:19:45.234825Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T21:19:37.456187Z digest=sha256:309d00754d73dde32a956ef5e5a65cd753a2f88f904a66f431dcdf24dcb9dc47

Observation d2bfe77f-6232-485e-94f8-8b69dbbaa289 · outbound

This paper cites Iterative Reasoning Preference Optimization.

From Answers to Rationales: Self-Aligning Multimodal Reasoning with Answer-Oriented Chain-of-Thought Iterative Reasoning Preference Optimization

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-06T21:19:37.514013Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:19:37.514013Z digest=sha256:a844ff16533bf521bd90e3ca498009e08d3010becb052ec9d2b3f1cc15c91fa1

Observation 1160b51d-4ce0-42a1-a8d0-d48d17eccc78 · outbound

This paper cites arXiv preprint arXiv:2405.17220 (2024).

From Answers to Rationales: Self-Aligning Multimodal Reasoning with Answer-Oriented Chain-of-Thought arXiv preprint arXiv:2405.17220 (2024)

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-06T21:19:37.569087Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:19:37.569087Z digest=sha256:7ae38c1394ea70a15268c7146c08052d45b42fc0ce7bce3ac3bc2e7ee9233654

Observation 064b4759-aea7-4fb9-ab43-37bc3456dccd · outbound

This paper cites Calibrated Self-Rewarding Vision Language Models.

From Answers to Rationales: Self-Aligning Multimodal Reasoning with Answer-Oriented Chain-of-Thought Calibrated Self-Rewarding Vision Language Models

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-06T21:19:37.631397Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:19:37.631397Z digest=sha256:45362e0eb8e45daaf0e2f78c241f292a30fa81e6a4701a4c59ede226d9c14130

Observation 38fbfaff-8ca6-4452-b784-652d4f837eca · outbound

This paper cites ACM MM (2024).

From Answers to Rationales: Self-Aligning Multimodal Reasoning with Answer-Oriented Chain-of-Thought ACM MM (2024)

Reference 55

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:19:45.222385Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T21:19:37.684578Z digest=sha256:fd0fca3801cf95f45d6b55b9ae94c425424a5ba6f7585f0ddb63ba20f218e1d1

Observation 7c2861a9-00e8-44db-8dbe-f02ff5886416 · outbound

This paper cites Math-LLaVA: Bootstrapping Mathematical Reasoning for Multimodal Large Language Models.

From Answers to Rationales: Self-Aligning Multimodal Reasoning with Answer-Oriented Chain-of-Thought Math-LLaVA: Bootstrapping Mathematical Reasoning for Multimodal Large Language Models

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-06T21:19:37.737245Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:19:37.737245Z digest=sha256:4743a45052a726d1d0aa5fb38889db1a3b814b193da3ce9f6e50129d6926f8ef

Observation df92a167-3d00-4554-abbe-a401f8be36b3 · outbound

This paper cites In: Findings of the Association for Computational Linguistics: ACL 2022, pp.

From Answers to Rationales: Self-Aligning Multimodal Reasoning with Answer-Oriented Chain-of-Thought In: Findings of the Association for Computational Linguistics: ACL 2022, pp

Reference 57

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:19:45.212095Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T21:19:37.787387Z digest=sha256:440bfd92bc34a91aad27feb7807d98f282634043f6ea0c0e7beb669b6651b487

Observation a6837e4b-3184-4146-8a75-ec98fa7cac91 · outbound

This paper cites an unresolved cited work.

From Answers to Rationales: Self-Aligning Multimodal Reasoning with Answer-Oriented Chain-of-Thought Unresolved cited work

Reference 58

Resolution
unresolved
raw_fallback, observed 2026-08-06T21:19:45.200325Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T21:19:37.847899Z digest=sha256:6370c2d4a0c2a7096bc46899c3ee70e508e8e3ebaed87612ffb4bbad43dc6a55

Observation 76c29561-5953-47e2-9bd0-8e5a3f1e83af · outbound

This paper cites NeurIPS (2022).

From Answers to Rationales: Self-Aligning Multimodal Reasoning with Answer-Oriented Chain-of-Thought NeurIPS (2022)

Reference 59

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:19:45.187971Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T21:19:37.904281Z digest=sha256:05cf840dee2cfa5a8c2fee3a573772bebd10d42b6bf3021b1bfe029914c9f345

Observation 506ad1ae-9ee1-4c99-bd5d-de37afc5029a · outbound

This paper cites Advances in neural information processing systems 33, 6840–6851 (2020).

From Answers to Rationales: Self-Aligning Multimodal Reasoning with Answer-Oriented Chain-of-Thought Advances in neural information processing systems 33, 6840–6851 (2020)

Reference 60

Resolution
unresolved
no resolver link, observed 2026-08-06T21:19:37.978169Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:19:37.978169Z digest=sha256:4cfce7632261fc411ac4cf0ce1456043bd486035bf7e6f2c45dfd7d6293edd59

Observation 902e720f-9e47-4ccb-9474-628f406e1f2c · outbound

This paper cites In: Findings of the Association for Computational Linguistics: EMNLP 2024, pp.

From Answers to Rationales: Self-Aligning Multimodal Reasoning with Answer-Oriented Chain-of-Thought In: Findings of the Association for Computational Linguistics: EMNLP 2024, pp

Reference 61

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:19:45.169529Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T21:19:38.078828Z digest=sha256:04d993d0a242c45767320beb7dfb64ef2bc98796836211a053fb5fa5e741f949

Observation b62d63ba-ff73-417d-8a51-463c0431f0e5 · outbound

This paper cites In: Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern 19 Recognition, pp.

From Answers to Rationales: Self-Aligning Multimodal Reasoning with Answer-Oriented Chain-of-Thought In: Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern 19 Recognition, pp

Reference 62

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:19:45.159105Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T21:19:38.187584Z digest=sha256:eab65a0239dfef53d7fe0f4a9146367c9059378be7d10ce2ce2cf6277575a658

Observation b747dfa7-a7fa-4294-97d0-637d11820d2a · outbound

This paper cites Qwen-VL: A Versatile Vision-Language Model for Understanding, Localization, Text Reading, and Beyond.

From Answers to Rationales: Self-Aligning Multimodal Reasoning with Answer-Oriented Chain-of-Thought Qwen-VL: A Versatile Vision-Language Model for Understanding, Localization, Text Reading, and Beyond

Reference 63

Resolution
unresolved
no resolver link, observed 2026-08-06T21:19:38.241871Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:19:38.241871Z digest=sha256:c65a0ddbda127f537cdb9f4216506c108a64ef5b29662906c096d9032c6cc7c7

Observation 363b3e55-d3a6-48f2-a8fb-aa67c8c7ff9f · outbound

This paper cites Gemini: A Family of Highly Capable Multimodal Models.

From Answers to Rationales: Self-Aligning Multimodal Reasoning with Answer-Oriented Chain-of-Thought Gemini: A Family of Highly Capable Multimodal Models

Reference 64

Resolution
unresolved
no resolver link, observed 2026-08-06T21:19:38.244912Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:19:38.244912Z digest=sha256:3d071ebed3a03aa122ea2464b3ee6969bad7be6a84905efcded7ab169b8bd138

Observation e61e0aa2-a79c-49ca-a8a6-0e3d3924ef27 · outbound

This paper cites Mixed Preference Optimization: Reinforcement Learning with Data Selection and Better Reference Model.

From Answers to Rationales: Self-Aligning Multimodal Reasoning with Answer-Oriented Chain-of-Thought Mixed Preference Optimization: Reinforcement Learning with Data Selection and Better Reference Model

Reference 65

Resolution
unresolved
no resolver link, observed 2026-08-06T21:19:38.248448Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:19:38.248448Z digest=sha256:4155ba45ae535099ff63ebc105e91e20d8546c95baa4f092aa0aaab5089b2cf2

Observation 27597d7a-d8a5-40f7-a23d-2c0e71ada57e · outbound

This paper cites GPT-4 Technical Report.

From Answers to Rationales: Self-Aligning Multimodal Reasoning with Answer-Oriented Chain-of-Thought GPT-4 Technical Report

Reference 66

Resolution
unresolved
no resolver link, observed 2026-08-06T21:19:38.285085Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:19:38.285085Z digest=sha256:6befd36bf561936b04f4103ad22dcf2e3dd69f30a175f66d4084f3bf81ffe7d6

Observation d137d441-3785-40b9-8f92-a61768b12a98 · outbound

This paper cites The Llama 3 Herd of Models.

From Answers to Rationales: Self-Aligning Multimodal Reasoning with Answer-Oriented Chain-of-Thought The Llama 3 Herd of Models

Reference 67

Resolution
unresolved
no resolver link, observed 2026-08-06T21:19:38.412010Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:19:38.412010Z digest=sha256:37dbbfc6ace43542e7e579877c443e9a07165ed1dca164e07049fd98ecdb5063

Observation b7f589be-bd18-4eec-babf-f731bf6cb5c3 · outbound

This paper cites Sub-answers:.

From Answers to Rationales: Self-Aligning Multimodal Reasoning with Answer-Oriented Chain-of-Thought Sub-answers:

Reference 68

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:19:45.146448Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T21:19:38.524221Z digest=sha256:dd5e5cfb89e44ea89ba23b0d3c5cd5d9f5dc12db2f7581e6593174a6a37db153

Observation 30d283c8-525d-4455-9304-d74c5c41f5a0 · outbound

This paper cites an unresolved cited work.

From Answers to Rationales: Self-Aligning Multimodal Reasoning with Answer-Oriented Chain-of-Thought Unresolved cited work

Reference 69

Resolution
unresolved
raw_fallback, observed 2026-08-06T21:19:45.134735Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T21:19:38.640246Z digest=sha256:750545902e0d3d1258a968ca9bf5a40e3ae7d6b2ad98a00c42d85b23ef719454

Observation e81b333b-9878-4b08-9774-621548c8a829 · outbound

This paper cites Uncertain.

From Answers to Rationales: Self-Aligning Multimodal Reasoning with Answer-Oriented Chain-of-Thought Uncertain

Reference 70

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:19:45.123412Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T21:19:38.712756Z digest=sha256:bf1081c16828c1d347a7ccc1bd997b7ee966a34ff3eab72f54261a902ef25c33

Observation ff08c9b2-35bd-46ab-aee7-14311f515173 · outbound

This paper cites an unresolved cited work.

From Answers to Rationales: Self-Aligning Multimodal Reasoning with Answer-Oriented Chain-of-Thought Unresolved cited work

Reference 71

Resolution
unresolved
raw_fallback, observed 2026-08-06T21:19:45.111855Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T21:19:38.821657Z digest=sha256:f8b652eb402651cb9c7f053ec350690e69a672d3b9661b800010e9b190745c29

Observation 28c47846-b7a7-4869-bd4c-9f3c93f91451 · outbound

This paper cites an unresolved cited work.

From Answers to Rationales: Self-Aligning Multimodal Reasoning with Answer-Oriented Chain-of-Thought Unresolved cited work

Reference 72

Resolution
unresolved
raw_fallback, observed 2026-08-06T21:19:45.101612Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T21:19:38.986073Z digest=sha256:15ca24675e253860ea9591ab00541496e0539e152c69829a587e130433c874da

Observation 0934a1a5-231d-4d5a-84a4-3258eaca149e · outbound

This paper cites The formula for the circumference C of a circle is given by: C = 2πr, where r is the radius of the circle.

From Answers to Rationales: Self-Aligning Multimodal Reasoning with Answer-Oriented Chain-of-Thought The formula for the circumference C of a circle is given by: C = 2πr, where r is the radius of the circle

Reference 73

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:19:45.090837Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T21:19:39.149496Z digest=sha256:341fe246e56f4a42146a97502d045624c805642b0a04c20fe3a79a66cc0d1fef

Observation 21841584-672d-4710-bfc0-f5a8b1dbf116 · outbound

This paper cites - Each side of the square is equal to the height of the rectangle, which is given as 32 units.

From Answers to Rationales: Self-Aligning Multimodal Reasoning with Answer-Oriented Chain-of-Thought - Each side of the square is equal to the height of the rectangle, which is given as 32 units

Reference 74

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:19:45.078402Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T21:19:39.353941Z digest=sha256:03bab7db59baa190616671f76be50d94cf492e9fadf5c2c2c678bfc60a2de103

Observation fe9dc88b-a511-49e8-b5a3-7bb538349296 · outbound

This paper cites - The diagonal of a square with side length s is given by s\sqrt{2}.

From Answers to Rationales: Self-Aligning Multimodal Reasoning with Answer-Oriented Chain-of-Thought - The diagonal of a square with side length s is given by s\sqrt{2}

Reference 75

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:19:45.066148Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T21:19:39.555704Z digest=sha256:28b7f009410a128564039858b07cbe82f41c796b0099eeb57d09c2679825dfb1

Observation 75072758-ec69-4e69-ae37-13309c87ea6b · outbound

This paper cites objects": [ {.

From Answers to Rationales: Self-Aligning Multimodal Reasoning with Answer-Oriented Chain-of-Thought objects": [ {

Reference 76

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:19:45.055324Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T21:19:39.697546Z digest=sha256:0c7f2746b363516813734f1a83e96767989c8dfebaa4dae3222a7de016c31e6b

Observation ce7acc8a-7b65-4e0e-a640-cca607d7fb03 · outbound

This paper cites an unresolved cited work.

From Answers to Rationales: Self-Aligning Multimodal Reasoning with Answer-Oriented Chain-of-Thought Unresolved cited work

Reference 77

Resolution
unresolved
raw_fallback, observed 2026-08-06T21:19:45.045568Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T21:19:39.904367Z digest=sha256:4b9812156e7c7ccd3b352b2922fe4864568f3e03c07f5243175a2af1384d2b03

Observation 7b657692-e593-4303-be7e-ab141c48ca68 · outbound

This paper cites objects": [ {.

From Answers to Rationales: Self-Aligning Multimodal Reasoning with Answer-Oriented Chain-of-Thought objects": [ {

Reference 78

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:19:45.036316Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T21:19:40.066958Z digest=sha256:dcb55b5d6be48ea7dbd9547ac6cee602f23f48c7d2822650e69d8b6e4eecb0c6

Observation ee271c14-4e87-4c91-bd70-311482409aea · outbound

This paper cites an unresolved cited work.

From Answers to Rationales: Self-Aligning Multimodal Reasoning with Answer-Oriented Chain-of-Thought Unresolved cited work

Reference 79

Resolution
unresolved
raw_fallback, observed 2026-08-06T21:19:45.026104Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T21:19:40.197225Z digest=sha256:6ed74909a23b25cd7643559ea849869dcae90fa55c4706a31bdbd9165619fe9e

Observation f88a98fc-b701-4a4d-bfd9-f731d866947b · outbound

This paper cites an unresolved cited work.

From Answers to Rationales: Self-Aligning Multimodal Reasoning with Answer-Oriented Chain-of-Thought Unresolved cited work

Reference 80

Resolution
unresolved
raw_fallback, observed 2026-08-06T21:19:45.016416Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T21:19:40.353407Z digest=sha256:b5a83a991ed674a301e0e693c79f179dccbe35e27b8d592018523657fc8fb28e

Observation 1883a4f1-7c59-4490-bf2f-b2cf1a354f1c · outbound

This paper cites an unresolved cited work.

From Answers to Rationales: Self-Aligning Multimodal Reasoning with Answer-Oriented Chain-of-Thought Unresolved cited work

Reference 81

Resolution
unresolved
raw_fallback, observed 2026-08-06T21:19:45.007140Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T21:19:40.505961Z digest=sha256:d285315f34f8f3374550a899d1aba83069cc689338c5aa2170db90ba2b3a6053

Observation 2ad79d24-12a9-441e-91b7-affaaf63f698 · outbound

This paper cites an unresolved cited work.

From Answers to Rationales: Self-Aligning Multimodal Reasoning with Answer-Oriented Chain-of-Thought Unresolved cited work

Reference 82

Resolution
unresolved
raw_fallback, observed 2026-08-06T21:19:44.997450Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T21:19:40.674246Z digest=sha256:bd2b8ea7173056de2ca2687471a0af8ecd1a14572ee0f7d45b4c2ee1136285c2

Observation d0b55959-a5ab-490a-98d2-0e3beca719df · outbound

This paper cites an unresolved cited work.

From Answers to Rationales: Self-Aligning Multimodal Reasoning with Answer-Oriented Chain-of-Thought Unresolved cited work

Reference 83

Resolution
unresolved
raw_fallback, observed 2026-08-06T21:19:44.988969Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T21:19:40.913540Z digest=sha256:c767fcac41812d48f8d1605737f490ddc3f324465d207b0db8bd8f92e55a4b39

Observation c222e408-bde4-4718-89c1-13c4082e2752 · outbound

This paper cites an unresolved cited work.

From Answers to Rationales: Self-Aligning Multimodal Reasoning with Answer-Oriented Chain-of-Thought Unresolved cited work

Reference 84

Resolution
unresolved
raw_fallback, observed 2026-08-06T21:19:44.981093Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T21:19:41.053299Z digest=sha256:39182374165adb327c718f241aa2080ce49a0074b32238d17754ce031b11ceda

Observation e7512077-b686-4bd8-a9b3-5cea7199c551 · outbound

This paper cites an unresolved cited work.

From Answers to Rationales: Self-Aligning Multimodal Reasoning with Answer-Oriented Chain-of-Thought Unresolved cited work

Reference 85

Resolution
unresolved
raw_fallback, observed 2026-08-06T21:19:44.972431Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T21:19:41.231965Z digest=sha256:2899dc1f23df2424cc6869d5a03a257a9f6674463f265c4f40fe8cdc6951c12a

Observation 56bc3bf0-545e-492a-9f06-abd4b9e3ffb5 · outbound

This paper cites objects": [{.

From Answers to Rationales: Self-Aligning Multimodal Reasoning with Answer-Oriented Chain-of-Thought objects": [{

Reference 86

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:19:44.964871Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T21:19:41.411657Z digest=sha256:4563751104adf666d4a4d316ce7cbbae960fc1508dab19366b755ee756cf07b9

Observation f684a584-5991-4a67-a1aa-b443994bcc49 · outbound

This paper cites **E** - Early blastocyst 3.

From Answers to Rationales: Self-Aligning Multimodal Reasoning with Answer-Oriented Chain-of-Thought **E** - Early blastocyst 3

Reference 87

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:19:44.957213Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T21:19:41.562213Z digest=sha256:1e1b7ce43809d442f5658d3afc0b1d2887b9e67dc400ea21864b45d4a7966fbb

Observation 11dbcc7f-07d4-4bfd-9fd0-89dcdfd5d89f · outbound

This paper cites (D) No.

From Answers to Rationales: Self-Aligning Multimodal Reasoning with Answer-Oriented Chain-of-Thought (D) No

Reference 88

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:19:44.948916Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T21:19:41.753763Z digest=sha256:d36e083a52f1c5ff942526772195627c3f993ef4274b6c737eb1ae3d5d4bdd2f

Observation 7448bd87-3e1f-46a6-9bd6-fc8b96006fee · outbound

This paper cites an unresolved cited work.

From Answers to Rationales: Self-Aligning Multimodal Reasoning with Answer-Oriented Chain-of-Thought Unresolved cited work

Reference 89

Resolution
unresolved
raw_fallback, observed 2026-08-06T21:19:44.939965Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T21:19:41.888306Z digest=sha256:b2b474507592353a76b187bb5e3b8f6763454649ae5e4c536afd4d6e5bec7913

Observation fc7a62d8-836e-4b9c-85a4-3c8aa8f6a1cd · outbound

This paper cites an unresolved cited work.

From Answers to Rationales: Self-Aligning Multimodal Reasoning with Answer-Oriented Chain-of-Thought Unresolved cited work

Reference 90

Resolution
unresolved
raw_fallback, observed 2026-08-06T21:19:44.930051Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T21:19:42.048251Z digest=sha256:f09c68ad00535493cbdd4363e57f9185ac22f73519fb2f0e366f9363496369c2

Observation e6b28067-6d4a-4129-a304-bffe378bbf2b · outbound

This paper cites an unresolved cited work.

From Answers to Rationales: Self-Aligning Multimodal Reasoning with Answer-Oriented Chain-of-Thought Unresolved cited work

Reference 91

Resolution
unresolved
raw_fallback, observed 2026-08-06T21:19:44.920005Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T21:19:42.250601Z digest=sha256:19d2ce3f376f8e1421057cacbce9880e961be0a482ccbc82f4f45e61525579c0

Observation 52c84aee-7ee7-4961-a5b7-29a97ddc5936 · outbound

This paper cites an unresolved cited work.

From Answers to Rationales: Self-Aligning Multimodal Reasoning with Answer-Oriented Chain-of-Thought Unresolved cited work

Reference 92

Resolution
unresolved
raw_fallback, observed 2026-08-06T21:19:44.850249Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T21:19:42.419828Z digest=sha256:f94663026d0356330a5571d8529dc987e523cc5772f3bce8e1656edf2165c958

Observation ab215d8c-386d-424d-8cac-6d20dfabc0da · outbound

This paper cites an unresolved cited work.

From Answers to Rationales: Self-Aligning Multimodal Reasoning with Answer-Oriented Chain-of-Thought Unresolved cited work

Reference 93

Resolution
unresolved
raw_fallback, observed 2026-08-06T21:19:44.638857Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T21:19:42.629321Z digest=sha256:0be26c767a0baac410c62f9feb7b69097bba57ca19929d26d606ce50fcc7c11b

Observation bdeacc47-2d6a-4866-ae5d-8a665d175dc6 · outbound

This paper cites objects": [ {.

From Answers to Rationales: Self-Aligning Multimodal Reasoning with Answer-Oriented Chain-of-Thought objects": [ {

Reference 94

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:19:44.429933Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T21:19:42.743404Z digest=sha256:a9898a08406b66808f58f6e7c10ad8dc1a744c67e3f458e2724114a476592fa7

Observation 35053a44-ae4e-4854-9b81-98b7c152a411 · outbound

This paper cites an unresolved cited work.

From Answers to Rationales: Self-Aligning Multimodal Reasoning with Answer-Oriented Chain-of-Thought Unresolved cited work

Reference 95

Resolution
unresolved
raw_fallback, observed 2026-08-06T21:19:44.238427Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T21:19:42.783517Z digest=sha256:88d1f8a08bab92a8a0635237b3e0ed9e8e251316b1cb0acf5402616ba14c7540

Observation 5b2c26af-e05c-445d-ad92-00dc4fc1c7ec · outbound

This paper cites an unresolved cited work.

From Answers to Rationales: Self-Aligning Multimodal Reasoning with Answer-Oriented Chain-of-Thought Unresolved cited work

Reference 96

Resolution
unresolved
raw_fallback, observed 2026-08-06T21:19:44.075489Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T21:19:42.836296Z digest=sha256:58769fb18cf25e3074ccbccde7b20daa019f1211f3f4fb96aa5faecbf08f51a0

Observation 0700ebef-9854-4246-8fec-40276333546d · outbound

This paper cites an unresolved cited work.

From Answers to Rationales: Self-Aligning Multimodal Reasoning with Answer-Oriented Chain-of-Thought Unresolved cited work

Reference 97

Resolution
unresolved
raw_fallback, observed 2026-08-06T21:19:43.777934Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T21:19:42.927117Z digest=sha256:f879144810ecb834eecb7cd4b8db3c5cf544c1c52c96644674a83bf2fccdd1bf

Observation 5da17c2f-9656-45aa-af6a-8b064383b604 · outbound

This paper cites The correct answer is: (C) smaller than Qwen-VL-7B + SMART : The derivative of the function y = log_2(x) is \frac{1}{x ln(2)}.

From Answers to Rationales: Self-Aligning Multimodal Reasoning with Answer-Oriented Chain-of-Thought The correct answer is: (C) smaller than Qwen-VL-7B + SMART : The derivative of the function y = log_2(x) is \frac{1}{x ln(2)}

Reference 98

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:19:43.630021Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T21:19:42.977063Z digest=sha256:d27f8562fb253a4346d37a10c3f6deaf7ed244369a91103683873003cfc9af8c

Pith citing papers

No inbound Pith citation observations are available.