Pith. sign in

Paper Citation Record · LEDGER

From Answers to Rationales: Self-Aligning Multimodal Reasoning with Answer-Oriented Chain-of-Thought

As of 7 August 2026, this Paper Citation Record lists 98 of 98 outbound references and 0 inbound Pith citation observations for arXiv:2507.02984.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2507.02984 v2

Coverage vector

measured 98 of 98 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-06T21:19:42.977063Z

measured 98 of 98 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-06T06:34:29.942622+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

98 of 98 outbound references displayed

  • verified exact0
  • verified fuzzy22
  • unresolved75
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch1

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation f416be21-c02a-4b9b-af2c-fcf74533376d · outbound

This paper cites MiniGPT-v2: large language model as a unified interface for vision-language multi-task learning.

From Answers to Rationales: Self-Aligning Multimodal Reasoning with Answer-Oriented Chain-of-Thought MiniGPT-v2: large language model as a unified interface for vision-language multi-task learning

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-06T21:19:33.852245Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:19:33.852245Z digest=sha256:ed3d4e63476fb4da5f7e761f0bba6c8fc9a8b969b5f9115b9ac9a53c69fb3180

Observation 6b577f96-c747-4113-88ce-7e575130aeed · outbound

This paper cites Shikra: Unleashing Multimodal LLM's Referential Dialogue Magic.

From Answers to Rationales: Self-Aligning Multimodal Reasoning with Answer-Oriented Chain-of-Thought Shikra: Unleashing Multimodal LLM's Referential Dialogue Magic

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-06T21:19:33.912791Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:19:33.912791Z digest=sha256:471d6f9a1f319714616ad7f867aad011194db1363cf6554817a24c8bdccbab21

Observation 13ec6268-f581-4eb7-a91e-dd40699eb75b · outbound

This paper cites Kosmos-2: Grounding Multimodal Large Language Models to the World.

From Answers to Rationales: Self-Aligning Multimodal Reasoning with Answer-Oriented Chain-of-Thought Kosmos-2: Grounding Multimodal Large Language Models to the World

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-06T21:19:34.013688Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:19:34.013688Z digest=sha256:2ab0d9d828e687c51fcb1c984edced9bfbfb8fd2931ee2d0653d46dcf429f792

Observation fe18d58a-759a-4aeb-a4ab-e676cca4ece0 · outbound

This paper cites CogVLM2: Visual Language Models for Image and Video Understanding.

From Answers to Rationales: Self-Aligning Multimodal Reasoning with Answer-Oriented Chain-of-Thought CogVLM2: Visual Language Models for Image and Video Understanding

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-06T21:19:34.077710Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:19:34.077710Z digest=sha256:6b439b981592f273537e3505eba3814f2604fed382775524271dbc9f784ec2b7

Observation 9f277b7f-a5e5-4af2-999b-fbd7e98ae66f · outbound

This paper cites In: International Conference on Machine Learning, pp.

From Answers to Rationales: Self-Aligning Multimodal Reasoning with Answer-Oriented Chain-of-Thought In: International Conference on Machine Learning, pp

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-06T21:19:34.177966Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:19:34.177966Z digest=sha256:41027ea9311fa362c9160ef005b101a9a35e0940b7f08a14aa31ea60753b0090

Observation 1f93425b-bfc9-47e1-a7d1-4d58c7ce2b8d · outbound

This paper cites In: Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition, pp.

From Answers to Rationales: Self-Aligning Multimodal Reasoning with Answer-Oriented Chain-of-Thought In: Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition, pp

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-06T21:19:34.243662Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:19:34.243662Z digest=sha256:6ed95ffc4835c36241348e3df4bef28e3f59f97efad63c32b854ece47729d491

Observation 9a18fa4d-9cd2-47c2-b3ad-e2f59471e96d · outbound

This paper cites In: Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition, pp.

From Answers to Rationales: Self-Aligning Multimodal Reasoning with Answer-Oriented Chain-of-Thought In: Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition, pp

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-06T21:19:34.353123Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:19:34.353123Z digest=sha256:16929cfe495d82391f746f353b87e36ddc712f9bbe967a4cc58161601cbe2c3c

Observation 927fd270-58f6-46c2-ade2-ded5b8f8ed54 · outbound

This paper cites SPHINX: The Joint Mixing of Weights, Tasks, and Visual Embeddings for Multi-modal Large Language Models.

From Answers to Rationales: Self-Aligning Multimodal Reasoning with Answer-Oriented Chain-of-Thought SPHINX: The Joint Mixing of Weights, Tasks, and Visual Embeddings for Multi-modal Large Language Models

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-06T21:19:34.422630Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:19:34.422630Z digest=sha256:0f48bcfa2c724dfbca641b4e318beab6b2207f421e9bf5af91266e5f2cdf0a1e

Observation 3d524dfb-0e36-4805-ba3e-b1bbf50ed920 · outbound

This paper cites How Far Are We to GPT-4V? Closing the Gap to Commercial Multimodal Models with Open-Source Suites.

From Answers to Rationales: Self-Aligning Multimodal Reasoning with Answer-Oriented Chain-of-Thought How Far Are We to GPT-4V? Closing the Gap to Commercial Multimodal Models with Open-Source Suites

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-06T21:19:34.505690Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:19:34.505690Z digest=sha256:4b255ca7c2ef977ccfbc9f0dcaa1de5b8d924afe7122f178bbbfaa881f2b88c5

Observation 6c9756ba-a469-4149-b741-3c25008706bd · outbound

This paper cites DeepSeek-VL: Towards Real-World Vision-Language Understanding.

From Answers to Rationales: Self-Aligning Multimodal Reasoning with Answer-Oriented Chain-of-Thought DeepSeek-VL: Towards Real-World Vision-Language Understanding

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-06T21:19:34.629464Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:19:34.629464Z digest=sha256:05128603182fa359bd981e8e053899f44d376ff2d59c435df0abaefb358bd0e1

Observation e3ec2494-dc05-4060-ae76-a6cb5f1f5a56 · outbound

This paper cites Yi: Open Foundation Models by 01.AI.

From Answers to Rationales: Self-Aligning Multimodal Reasoning with Answer-Oriented Chain-of-Thought Yi: Open Foundation Models by 01.AI

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-06T21:19:34.703051Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:19:34.703051Z digest=sha256:755d20785e6ba4ad50c448770007bd31498180f3254a02e579e085513f1efabb

Observation cc9022a1-7334-4aff-8f1d-9af90499af99 · outbound

This paper cites https://llava-vl.github.

From Answers to Rationales: Self-Aligning Multimodal Reasoning with Answer-Oriented Chain-of-Thought https://llava-vl.github

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-06T21:19:34.772600Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:19:34.772600Z digest=sha256:b306617bb9a3102e94c563024576121f1daca05d951673bf1f9a80729b36f52b

Observation 055f57a2-4f6f-4f21-b0e6-8ab138540039 · outbound

This paper cites MiniCPM-V: A GPT-4V Level MLLM on Your Phone.

From Answers to Rationales: Self-Aligning Multimodal Reasoning with Answer-Oriented Chain-of-Thought MiniCPM-V: A GPT-4V Level MLLM on Your Phone

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-06T21:19:34.859505Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:19:34.859505Z digest=sha256:b4a718fc24c71ca6783a860aa39ac4088ad1e9f18e365b904a9e939e2278a32c

Observation 2422e848-f067-473f-bb2f-4a07fd237b05 · outbound

This paper cites Qwen2-VL: Enhancing Vision-Language Model's Perception of the World at Any Resolution.

From Answers to Rationales: Self-Aligning Multimodal Reasoning with Answer-Oriented Chain-of-Thought Qwen2-VL: Enhancing Vision-Language Model's Perception of the World at Any Resolution

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-06T21:19:34.937712Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:19:34.937712Z digest=sha256:fef97ad598f436a89eea8c144567dc2a7013f235b69a9f90b2e46162e0a20b36

Observation 4708ab48-5e6b-485b-9155-412f188f8d23 · outbound

This paper cites AMBER: An LLM-free Multi-dimensional Benchmark for MLLMs Hallucination Evaluation.

From Answers to Rationales: Self-Aligning Multimodal Reasoning with Answer-Oriented Chain-of-Thought AMBER: An LLM-free Multi-dimensional Benchmark for MLLMs Hallucination Evaluation

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-06T21:19:35.022823Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:19:35.022823Z digest=sha256:bb4edf4a510b469126fa26f873fc33ec20a96b8761703a4e0a7ff3fe0e6bbf1c

Observation 7900ed98-6e53-4bac-8741-2e33725432fe · outbound

This paper cites NeurIPS (2023).

From Answers to Rationales: Self-Aligning Multimodal Reasoning with Answer-Oriented Chain-of-Thought NeurIPS (2023)

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-06T21:19:35.110223Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:19:35.110223Z digest=sha256:f152d4f0760ecbe5668061723470752f2933782f5b21665c406985480ce28b93

Observation b2ad41e5-2682-428e-b7ab-be0596cb1577 · outbound

This paper cites MME: A Comprehensive Evaluation Benchmark for Multimodal Large Language Models.

From Answers to Rationales: Self-Aligning Multimodal Reasoning with Answer-Oriented Chain-of-Thought MME: A Comprehensive Evaluation Benchmark for Multimodal Large Language Models

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-06T21:19:35.203253Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:19:35.203253Z digest=sha256:6a65992a191bd13f86ebfa15ade38fcde841a47634d2eb4c8cf9135dbd806157

Observation e4225a55-a0a2-487c-9e2e-1be11a7eea43 · outbound

This paper cites In: Proceedings of the IEEE International Conference on Computer Vision, pp.

From Answers to Rationales: Self-Aligning Multimodal Reasoning with Answer-Oriented Chain-of-Thought In: Proceedings of the IEEE International Conference on Computer Vision, pp

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-06T21:19:35.260301Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:19:35.260301Z digest=sha256:33343a29d8cfe5f8207bdd95fb16cbb6e318cda40e8b0b32e72832caaa84bad2

Observation 26b4a820-0ff1-4e90-b6c6-3dfd42f199ec · outbound

This paper cites In: Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition, pp.

From Answers to Rationales: Self-Aligning Multimodal Reasoning with Answer-Oriented Chain-of-Thought In: Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition, pp

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-06T21:19:35.355158Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:19:35.355158Z digest=sha256:69c4c8611d61623e52f166ab2afb46044e3204c29effac5d4cdf485947d64278

Observation 04819b31-00e6-4373-a1cb-084c8649815d · outbound

This paper cites MM-Vet: Evaluating Large Multimodal Models for Integrated Capabilities.

From Answers to Rationales: Self-Aligning Multimodal Reasoning with Answer-Oriented Chain-of-Thought MM-Vet: Evaluating Large Multimodal Models for Integrated Capabilities

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-06T21:19:35.419588Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:19:35.419588Z digest=sha256:3fc460e7d691fd55be8ea868216513b01a6860d66793329bdebcdb87a8703933

Observation 89997c7a-e103-4a9d-ad83-bba8518244d1 · outbound

This paper cites In: International Conference on Learning Representations (ICLR) (2024).

From Answers to Rationales: Self-Aligning Multimodal Reasoning with Answer-Oriented Chain-of-Thought In: International Conference on Learning Representations (ICLR) (2024)

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-06T21:19:35.492421Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:19:35.492421Z digest=sha256:567586fdd95f4b7ef05a96da87dbb9aea9047ce8ce7232aad9cc8e24811f247d

Observation 284e0a8b-399e-456d-b674-0fcf5fb39f21 · outbound

This paper cites ACL (2024).

From Answers to Rationales: Self-Aligning Multimodal Reasoning with Answer-Oriented Chain-of-Thought ACL (2024)

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-06T21:19:35.556816Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:19:35.556816Z digest=sha256:11de7a050148dcaff0b0a8e95199d307f3697b504c061c5afa86990426463091

Observation f07419c3-9f53-422c-82e6-d0314e8f14f2 · outbound

This paper cites NeurIPS (2022).

From Answers to Rationales: Self-Aligning Multimodal Reasoning with Answer-Oriented Chain-of-Thought NeurIPS (2022)

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-06T21:19:35.650695Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:19:35.650695Z digest=sha256:5fe0d562e563209fa1996f5b4b4f29ec0a775e2f231919e2fcb1d33270ad00f0

Observation d64192ec-5a48-479b-98c3-4c1ae57b6b90 · outbound

This paper cites Multimodal Chain-of-Thought Reasoning in Language Models.

From Answers to Rationales: Self-Aligning Multimodal Reasoning with Answer-Oriented Chain-of-Thought Multimodal Chain-of-Thought Reasoning in Language Models

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-06T21:19:35.695869Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:19:35.695869Z digest=sha256:e8a59f66d022b67bf276148de18dacadc762b1b905926e05602098005b311740

Observation 5dd3c81e-b9b3-49bf-8400-c58b711e5f56 · outbound

This paper cites NeurIPS (2023).

From Answers to Rationales: Self-Aligning Multimodal Reasoning with Answer-Oriented Chain-of-Thought NeurIPS (2023)

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-06T21:19:35.779999Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:19:35.779999Z digest=sha256:953113ddef58b35f017bb7061dc24ba5a6c1dfe098c763de422d118565d47fbc

Observation 089aa579-19a6-41ed-875b-9525f7e724b3 · outbound

This paper cites In: Proceedings of the AAAI Conference on Artificial Intelligence, vol.

From Answers to Rationales: Self-Aligning Multimodal Reasoning with Answer-Oriented Chain-of-Thought In: Proceedings of the AAAI Conference on Artificial Intelligence, vol

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-06T21:19:35.846251Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:19:35.846251Z digest=sha256:0067d0c6b7e602aa424fb7ad315bdacf187bd7d8c8e299c8b4b3abe1555d0987

Observation 9fe9cbd2-4458-4d38-98ff-8964cf14f0fc · outbound

This paper cites ACM MM (2024) 16.

From Answers to Rationales: Self-Aligning Multimodal Reasoning with Answer-Oriented Chain-of-Thought ACM MM (2024) 16

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-06T21:19:35.912123Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:19:35.912123Z digest=sha256:48adb5c5e12f3b104ee95c8baa6ce1fb627ca902d177492d80a695bca0dad07b

Observation 207edfa8-d3a2-4a54-a274-d4ba0697749a · outbound

This paper cites In: CVPR (2024).

From Answers to Rationales: Self-Aligning Multimodal Reasoning with Answer-Oriented Chain-of-Thought In: CVPR (2024)

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-06T21:19:35.992476Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:19:35.992476Z digest=sha256:9a83eadb57b66f94cd8a00d8e06bbea501fd12a126216092f29fc407efd7640c

Observation 518bea90-bf82-478e-9b14-1bafb206c194 · outbound

This paper cites The Role of Chain-of-Thought in Complex Vision-Language Reasoning Task.

From Answers to Rationales: Self-Aligning Multimodal Reasoning with Answer-Oriented Chain-of-Thought The Role of Chain-of-Thought in Complex Vision-Language Reasoning Task

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-06T21:19:36.092686Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:19:36.092686Z digest=sha256:a8507b3cf59f05af3f4a5ebd48f16ca5f6b85284a83ad374c1fd38d6aad3b280

Observation 5b546750-8f74-4020-b4b0-484e32544102 · outbound

This paper cites Enhancing Large Vision Language Models with Self-Training on Image Comprehension.

From Answers to Rationales: Self-Aligning Multimodal Reasoning with Answer-Oriented Chain-of-Thought Enhancing Large Vision Language Models with Self-Training on Image Comprehension

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-06T21:19:36.169281Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:19:36.169281Z digest=sha256:d3545ac204d47f20cb1112e68c1bb9cb38a569bfee50d14e1931488fe19d0af3

Observation dcc6de35-d0ab-40e1-9116-b0127168f179 · outbound

This paper cites ISR-DPO: Aligning Large Multimodal Models for Videos by Iterative Self-Retrospective DPO.

From Answers to Rationales: Self-Aligning Multimodal Reasoning with Answer-Oriented Chain-of-Thought ISR-DPO: Aligning Large Multimodal Models for Videos by Iterative Self-Retrospective DPO

Reference 31

Resolution
metadata mismatch
local_arxiv, observed 2026-08-06T21:19:43.391954Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-06T21:19:36.242869Z digest=sha256:070627ad15b0ce8667c472e84d09ae5f6653578639b40aaf14e3b494dd3aca1a

Observation 898c1883-d398-4829-a80a-277d07070c79 · outbound

This paper cites Enhancing Visual-Language Modality Alignment in Large Vision Language Models via Self-Improvement.

From Answers to Rationales: Self-Aligning Multimodal Reasoning with Answer-Oriented Chain-of-Thought Enhancing Visual-Language Modality Alignment in Large Vision Language Models via Self-Improvement

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-06T21:19:36.307866Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:19:36.307866Z digest=sha256:ce51c44fe72413204f1f8184b71b3f0b6ab6a7dcae1bcaf7eedc3b4b0f17f315

Observation f00da12f-f109-47a1-9460-1bdec9af2d54 · outbound

This paper cites Beyond Human Data: Aligning Multimodal Large Language Models by Iterative Self-Evolution.

From Answers to Rationales: Self-Aligning Multimodal Reasoning with Answer-Oriented Chain-of-Thought Beyond Human Data: Aligning Multimodal Large Language Models by Iterative Self-Evolution

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-06T21:19:36.373666Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:19:36.373666Z digest=sha256:2cfd45f6a174a7cb52e2ef879fed7765c68c8bea83277bda291a8d42e3a6275a

Observation c9fb9dea-dadb-44bc-9070-4029abaf468e · outbound

This paper cites In: Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition, pp.

From Answers to Rationales: Self-Aligning Multimodal Reasoning with Answer-Oriented Chain-of-Thought In: Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition, pp

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-06T21:19:36.451178Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:19:36.451178Z digest=sha256:6aba8e28c15e3309c0c54b50da6cb254ccc115ad292b2da8d609d652e7ae609b

Observation 56362f2f-5a59-4b3c-83fc-daa77b6a2a8b · outbound

This paper cites In: Computer Vision–ECCV 2016: 14th European Conference, Amsterdam, The Netherlands, October 11–14, 2016, Proceedings, Part IV 14, pp.

From Answers to Rationales: Self-Aligning Multimodal Reasoning with Answer-Oriented Chain-of-Thought In: Computer Vision–ECCV 2016: 14th European Conference, Amsterdam, The Netherlands, October 11–14, 2016, Proceedings, Part IV 14, pp

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-06T21:19:36.528716Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:19:36.528716Z digest=sha256:49eb9177c5a252f2b37460ee5dce79aba7f7dca1a1fec1177e20ca887588c334

Observation 18401eb3-51b1-4d45-bc32-17a14552b1ca · outbound

This paper cites an unresolved cited work.

From Answers to Rationales: Self-Aligning Multimodal Reasoning with Answer-Oriented Chain-of-Thought Unresolved cited work

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-06T21:19:36.608901Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:19:36.608901Z digest=sha256:05b62dbb23af06fa4fe29df6c016e43fb7cce4c88279064cd8d5cecf1788dcae

Observation e0aec378-e836-4353-a392-163c639c9984 · outbound

This paper cites Measuring Mathematical Problem Solving With the MATH Dataset.

From Answers to Rationales: Self-Aligning Multimodal Reasoning with Answer-Oriented Chain-of-Thought Measuring Mathematical Problem Solving With the MATH Dataset

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-06T21:19:36.653815Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:19:36.653815Z digest=sha256:3d8954453bba09cf6bc6a1ecaac15597fe4dbb6a213cb6b4479647417042b373

Observation b1d80a42-65fc-418e-a042-88622768a4dd · outbound

This paper cites In: European Conference on Computer Vision, pp.

From Answers to Rationales: Self-Aligning Multimodal Reasoning with Answer-Oriented Chain-of-Thought In: European Conference on Computer Vision, pp

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-06T21:19:36.720980Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:19:36.720980Z digest=sha256:7c2a1ff31482f5fef013456302b70f421fcab3e2f812eeedb5223e20d1a3f339

Observation 90347a6c-eb5a-4ac5-9b91-3134c0d0cba6 · outbound

This paper cites CoCoT: Contrastive Chain-of-Thought Prompting for Large Multimodal Models with Multiple Image Inputs.

From Answers to Rationales: Self-Aligning Multimodal Reasoning with Answer-Oriented Chain-of-Thought CoCoT: Contrastive Chain-of-Thought Prompting for Large Multimodal Models with Multiple Image Inputs

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-06T21:19:36.765868Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:19:36.765868Z digest=sha256:ee3c79147ce69844f13486bb61ca3f6d05afbf3a57723e62cf5693a4531b57b0

Observation ce0619b4-d8bf-4350-9aa7-441eacbc1a92 · outbound

This paper cites Visual CoT: Advancing Multi-Modal Language Models with a Comprehensive Dataset and Benchmark for Chain-of-Thought Reasoning.

From Answers to Rationales: Self-Aligning Multimodal Reasoning with Answer-Oriented Chain-of-Thought Visual CoT: Advancing Multi-Modal Language Models with a Comprehensive Dataset and Benchmark for Chain-of-Thought Reasoning

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-06T21:19:36.818352Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:19:36.818352Z digest=sha256:fc8269981466f6445df13f3a37188a4095dfb2eaad0f23fcfac0a78a188fb75e

Observation 488629da-b6f5-4919-b848-b543690d8c37 · outbound

This paper cites Journal of machine learning research 21(140), 1–67 (2020).

From Answers to Rationales: Self-Aligning Multimodal Reasoning with Answer-Oriented Chain-of-Thought Journal of machine learning research 21(140), 1–67 (2020)

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:19:45.279871Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-06T21:19:36.869481Z digest=sha256:9940411cf5084e17259a3e5af2eddc80de5148caf57bab9ea073ea6908043c1f

Observation a9ac94fc-c90b-4f57-b292-963270d15c1e · outbound

This paper cites : Training language models to follow instructions with human feedback.

From Answers to Rationales: Self-Aligning Multimodal Reasoning with Answer-Oriented Chain-of-Thought : Training language models to follow instructions with human feedback

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:19:45.268677Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-06T21:19:36.945519Z digest=sha256:ac3c7d8bea1b70d58ab7078bbfa67ddba0bbb7919470825ef796db860f704e33

Observation 8ad19dd6-f942-4ace-90d9-dab0c91f6bc5 · outbound

This paper cites LLaVA-NeXT-Interleave: Tackling Multi-image, Video, and 3D in Large Multimodal Models.

From Answers to Rationales: Self-Aligning Multimodal Reasoning with Answer-Oriented Chain-of-Thought LLaVA-NeXT-Interleave: Tackling Multi-image, Video, and 3D in Large Multimodal Models

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-06T21:19:36.992645Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:19:36.992645Z digest=sha256:c44a830bbf0db1c6b1d2bf549c950f4cd91147d65d519f07d3c291e941230a42

Observation 009094eb-d556-4465-a87d-d4e5bc3d9846 · outbound

This paper cites In: Proceedings of the 37th International Conference on Neural Information Processing Systems (2024).

From Answers to Rationales: Self-Aligning Multimodal Reasoning with Answer-Oriented Chain-of-Thought In: Proceedings of the 37th International Conference on Neural Information Processing Systems (2024)

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:19:45.257314Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-06T21:19:37.046109Z digest=sha256:718fc8bde05c916bb95f8c46e2064b01af9a7a11023f329245c6ad15224a7915

Observation a56f1546-e8ca-4dee-91e8-64d5b58d6023 · outbound

This paper cites https://openai.com/research/ gpt-4v-system-card.

From Answers to Rationales: Self-Aligning Multimodal Reasoning with Answer-Oriented Chain-of-Thought https://openai.com/research/ gpt-4v-system-card

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:19:45.245815Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-06T21:19:37.124103Z digest=sha256:5cc7247b078b11fed56ce190e283cba2e2fdd56e323e93df587e29f8f15527de

Observation 1f54acb1-611e-49c7-ba95-7ebfe0a7f823 · outbound

This paper cites Direct Language Model Alignment from Online AI Feedback.

From Answers to Rationales: Self-Aligning Multimodal Reasoning with Answer-Oriented Chain-of-Thought Direct Language Model Alignment from Online AI Feedback

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-06T21:19:37.182865Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:19:37.182865Z digest=sha256:4d4426f99aa1ee5ef66a254b4c4d96a27dde6bf22502f15d1610ac1977adbd6c

Observation 7c78759c-5123-4548-b72a-5f8e56dc45c4 · outbound

This paper cites Self-Rewarding Language Models.

From Answers to Rationales: Self-Aligning Multimodal Reasoning with Answer-Oriented Chain-of-Thought Self-Rewarding Language Models

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-06T21:19:37.231697Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:19:37.231697Z digest=sha256:790f075e1e7e1f78b727f20fd4cb9ea8055077b10f1c39797f08157760daeda7

Observation 15b86385-82d0-4368-b077-1c142b2587e1 · outbound

This paper cites Human Alignment of Large Language Models through Online Preference Optimisation.

From Answers to Rationales: Self-Aligning Multimodal Reasoning with Answer-Oriented Chain-of-Thought Human Alignment of Large Language Models through Online Preference Optimisation

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-06T21:19:37.284453Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:19:37.284453Z digest=sha256:2701a10d7f38a58a953e411c654dd1b15117252605c44b7cb517d2535a646ff0

Observation 67e203a4-c003-42bd-a4d7-b46b6906879b · outbound

This paper cites RLHF Workflow: From Reward Modeling to Online RLHF.

From Answers to Rationales: Self-Aligning Multimodal Reasoning with Answer-Oriented Chain-of-Thought RLHF Workflow: From Reward Modeling to Online RLHF

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-06T21:19:37.331130Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:19:37.331130Z digest=sha256:d6925bb579a722b13d9d99f361108898e62cb5ad760189ffc8ba4aac38c469ab

Observation 884b4a97-d57d-43a0-98d2-87d6e9ac026f · outbound

This paper cites Self-Play Fine-Tuning Converts Weak Language Models to Strong Language Models.

From Answers to Rationales: Self-Aligning Multimodal Reasoning with Answer-Oriented Chain-of-Thought Self-Play Fine-Tuning Converts Weak Language Models to Strong Language Models

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-06T21:19:37.398160Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:19:37.398160Z digest=sha256:8d212d8e074f6f1a38a05733686b1f779c58b5c78b3ceed11ab8d87dd6d08406

Observation 07f27bc1-2ecd-417a-b148-5556b19f72db · outbound

This paper cites Advances in Neural Information Processing Systems 35, 15476–15488 (2022).

From Answers to Rationales: Self-Aligning Multimodal Reasoning with Answer-Oriented Chain-of-Thought Advances in Neural Information Processing Systems 35, 15476–15488 (2022)

Reference 51

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:19:45.234825Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-06T21:19:37.456187Z digest=sha256:65e570478c722d458aae58fbfe3b27464f000876f631df4098ddedc2fb824ff9

Observation d2bfe77f-6232-485e-94f8-8b69dbbaa289 · outbound

This paper cites Iterative Reasoning Preference Optimization.

From Answers to Rationales: Self-Aligning Multimodal Reasoning with Answer-Oriented Chain-of-Thought Iterative Reasoning Preference Optimization

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-06T21:19:37.514013Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:19:37.514013Z digest=sha256:6e08b7f36ac266297bb6210d2920a97e7d77ce37391866035390ebaf9e695263

Observation 1160b51d-4ce0-42a1-a8d0-d48d17eccc78 · outbound

This paper cites arXiv preprint arXiv:2405.17220 (2024).

From Answers to Rationales: Self-Aligning Multimodal Reasoning with Answer-Oriented Chain-of-Thought arXiv preprint arXiv:2405.17220 (2024)

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-06T21:19:37.569087Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:19:37.569087Z digest=sha256:810de2226ce9807f41570608910f21f6606aaf5163bf11e9a61612f10b180aa5

Observation 064b4759-aea7-4fb9-ab43-37bc3456dccd · outbound

This paper cites Calibrated Self-Rewarding Vision Language Models.

From Answers to Rationales: Self-Aligning Multimodal Reasoning with Answer-Oriented Chain-of-Thought Calibrated Self-Rewarding Vision Language Models

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-06T21:19:37.631397Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:19:37.631397Z digest=sha256:cf3b79b45bb1202718b08ade7d0ca52e65abdd01e43425e37134e19ee2cb9c32

Observation 38fbfaff-8ca6-4452-b784-652d4f837eca · outbound

This paper cites ACM MM (2024).

From Answers to Rationales: Self-Aligning Multimodal Reasoning with Answer-Oriented Chain-of-Thought ACM MM (2024)

Reference 55

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:19:45.222385Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-06T21:19:37.684578Z digest=sha256:82e8c52a09acc582d7085c0b4c5243df2e4eb3019d4f9cbae6b1e1a6f8720446

Observation 7c2861a9-00e8-44db-8dbe-f02ff5886416 · outbound

This paper cites Math-LLaVA: Bootstrapping Mathematical Reasoning for Multimodal Large Language Models.

From Answers to Rationales: Self-Aligning Multimodal Reasoning with Answer-Oriented Chain-of-Thought Math-LLaVA: Bootstrapping Mathematical Reasoning for Multimodal Large Language Models

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-06T21:19:37.737245Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:19:37.737245Z digest=sha256:ab3e0dff0790a78712bcbe81172ccad1a70213b296eb1038a4fea600276d844a

Observation df92a167-3d00-4554-abbe-a401f8be36b3 · outbound

This paper cites In: Findings of the Association for Computational Linguistics: ACL 2022, pp.

From Answers to Rationales: Self-Aligning Multimodal Reasoning with Answer-Oriented Chain-of-Thought In: Findings of the Association for Computational Linguistics: ACL 2022, pp

Reference 57

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:19:45.212095Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-06T21:19:37.787387Z digest=sha256:5f9206ebcc4d8b630c34e8b2ba1ebea046caaff9ba6d1a305c927c0ee5269c48

Observation a6837e4b-3184-4146-8a75-ec98fa7cac91 · outbound

This paper cites an unresolved cited work.

From Answers to Rationales: Self-Aligning Multimodal Reasoning with Answer-Oriented Chain-of-Thought Unresolved cited work

Reference 58

Resolution
unresolved
raw_fallback, observed 2026-08-06T21:19:45.200325Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-06T21:19:37.847899Z digest=sha256:33270d231e76e5e16bee424393199f125f5e5c0648da6675f8f756ee3dd9abdd

Observation 76c29561-5953-47e2-9bd0-8e5a3f1e83af · outbound

This paper cites NeurIPS (2022).

From Answers to Rationales: Self-Aligning Multimodal Reasoning with Answer-Oriented Chain-of-Thought NeurIPS (2022)

Reference 59

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:19:45.187971Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-06T21:19:37.904281Z digest=sha256:564e491315a63c3bd142f63b83cbca013473e58eaa79aef1f18ad9d55360bb3b

Observation 506ad1ae-9ee1-4c99-bd5d-de37afc5029a · outbound

This paper cites Advances in neural information processing systems 33, 6840–6851 (2020).

From Answers to Rationales: Self-Aligning Multimodal Reasoning with Answer-Oriented Chain-of-Thought Advances in neural information processing systems 33, 6840–6851 (2020)

Reference 60

Resolution
unresolved
no resolver link, observed 2026-08-06T21:19:37.978169Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:19:37.978169Z digest=sha256:458534cd349cd65f59a7806c9392ac90f93cceedd905d3251c6087ea30a7f382

Observation 902e720f-9e47-4ccb-9474-628f406e1f2c · outbound

This paper cites In: Findings of the Association for Computational Linguistics: EMNLP 2024, pp.

From Answers to Rationales: Self-Aligning Multimodal Reasoning with Answer-Oriented Chain-of-Thought In: Findings of the Association for Computational Linguistics: EMNLP 2024, pp

Reference 61

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:19:45.169529Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-06T21:19:38.078828Z digest=sha256:20f1b12575e68bcb29433affc8ac01b1e6524357dc22ecec325a70c6fd7ab5f2

Observation b62d63ba-ff73-417d-8a51-463c0431f0e5 · outbound

This paper cites In: Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern 19 Recognition, pp.

From Answers to Rationales: Self-Aligning Multimodal Reasoning with Answer-Oriented Chain-of-Thought In: Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern 19 Recognition, pp

Reference 62

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:19:45.159105Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-06T21:19:38.187584Z digest=sha256:6c4d12d79cf0382e6881ef8c8d23faaea66a5c0f9087b15a46961d4e6351475e

Observation b747dfa7-a7fa-4294-97d0-637d11820d2a · outbound

This paper cites Qwen-VL: A Versatile Vision-Language Model for Understanding, Localization, Text Reading, and Beyond.

From Answers to Rationales: Self-Aligning Multimodal Reasoning with Answer-Oriented Chain-of-Thought Qwen-VL: A Versatile Vision-Language Model for Understanding, Localization, Text Reading, and Beyond

Reference 63

Resolution
unresolved
no resolver link, observed 2026-08-06T21:19:38.241871Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:19:38.241871Z digest=sha256:a897a2f7ff8b2b4a763bb534ea5f8e6c329011fbd35a24f222d88c146d8b3b07

Observation 363b3e55-d3a6-48f2-a8fb-aa67c8c7ff9f · outbound

This paper cites Gemini: A Family of Highly Capable Multimodal Models.

From Answers to Rationales: Self-Aligning Multimodal Reasoning with Answer-Oriented Chain-of-Thought Gemini: A Family of Highly Capable Multimodal Models

Reference 64

Resolution
unresolved
no resolver link, observed 2026-08-06T21:19:38.244912Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:19:38.244912Z digest=sha256:59942979ff073f31d87a586466170d05271a609809b9c0ef4811c99099d96591

Observation e61e0aa2-a79c-49ca-a8a6-0e3d3924ef27 · outbound

This paper cites Mixed Preference Optimization: Reinforcement Learning with Data Selection and Better Reference Model.

From Answers to Rationales: Self-Aligning Multimodal Reasoning with Answer-Oriented Chain-of-Thought Mixed Preference Optimization: Reinforcement Learning with Data Selection and Better Reference Model

Reference 65

Resolution
unresolved
no resolver link, observed 2026-08-06T21:19:38.248448Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:19:38.248448Z digest=sha256:ef37cc656ba44587ffc87a784934155e67e37be2b5ef520a94912c206c92eed7

Observation 27597d7a-d8a5-40f7-a23d-2c0e71ada57e · outbound

This paper cites GPT-4 Technical Report.

From Answers to Rationales: Self-Aligning Multimodal Reasoning with Answer-Oriented Chain-of-Thought GPT-4 Technical Report

Reference 66

Resolution
unresolved
no resolver link, observed 2026-08-06T21:19:38.285085Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:19:38.285085Z digest=sha256:2b2748a74228065266ea4842bfec43f2d4ed86d2cfa208e9229a2500250ce4b3

Observation d137d441-3785-40b9-8f92-a61768b12a98 · outbound

This paper cites The Llama 3 Herd of Models.

From Answers to Rationales: Self-Aligning Multimodal Reasoning with Answer-Oriented Chain-of-Thought The Llama 3 Herd of Models

Reference 67

Resolution
unresolved
no resolver link, observed 2026-08-06T21:19:38.412010Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:19:38.412010Z digest=sha256:2d2962e26969c65047dfd85642ef57a5534fed12380954ae13f272de4855eb76

Observation b7f589be-bd18-4eec-babf-f731bf6cb5c3 · outbound

This paper cites Sub-answers:.

From Answers to Rationales: Self-Aligning Multimodal Reasoning with Answer-Oriented Chain-of-Thought Sub-answers:

Reference 68

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:19:45.146448Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-06T21:19:38.524221Z digest=sha256:88827f9c5cadad9a0a57d7a5d440dc5e65e5aa1ddad2c06df06e7b64c8f9160f

Observation 30d283c8-525d-4455-9304-d74c5c41f5a0 · outbound

This paper cites an unresolved cited work.

From Answers to Rationales: Self-Aligning Multimodal Reasoning with Answer-Oriented Chain-of-Thought Unresolved cited work

Reference 69

Resolution
unresolved
raw_fallback, observed 2026-08-06T21:19:45.134735Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-06T21:19:38.640246Z digest=sha256:53715f757fc92999c77fa67d18dbbd790b2804c69c8a7f3ae40b8ad073088ead

Observation e81b333b-9878-4b08-9774-621548c8a829 · outbound

This paper cites Uncertain.

From Answers to Rationales: Self-Aligning Multimodal Reasoning with Answer-Oriented Chain-of-Thought Uncertain

Reference 70

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:19:45.123412Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-06T21:19:38.712756Z digest=sha256:7d2d24d1422f1b84bc968d654d921e24e112948ac84c8c9280758cc5100b3222

Observation ff08c9b2-35bd-46ab-aee7-14311f515173 · outbound

This paper cites an unresolved cited work.

From Answers to Rationales: Self-Aligning Multimodal Reasoning with Answer-Oriented Chain-of-Thought Unresolved cited work

Reference 71

Resolution
unresolved
raw_fallback, observed 2026-08-06T21:19:45.111855Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-06T21:19:38.821657Z digest=sha256:4845083b4f62e2d340724dc669651198546c91c8d7e3496cb352eab061bf17cb

Observation 28c47846-b7a7-4869-bd4c-9f3c93f91451 · outbound

This paper cites an unresolved cited work.

From Answers to Rationales: Self-Aligning Multimodal Reasoning with Answer-Oriented Chain-of-Thought Unresolved cited work

Reference 72

Resolution
unresolved
raw_fallback, observed 2026-08-06T21:19:45.101612Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-06T21:19:38.986073Z digest=sha256:c6050d875885f2de3b5d931a7023b8af88e6592b9937be349b7a8c532f7837e4

Observation 0934a1a5-231d-4d5a-84a4-3258eaca149e · outbound

This paper cites The formula for the circumference C of a circle is given by: C = 2πr, where r is the radius of the circle.

From Answers to Rationales: Self-Aligning Multimodal Reasoning with Answer-Oriented Chain-of-Thought The formula for the circumference C of a circle is given by: C = 2πr, where r is the radius of the circle

Reference 73

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:19:45.090837Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-06T21:19:39.149496Z digest=sha256:9e12544956d7394ff7b67067a29b2aeb5504ab5bae502305316e72a258872abe

Observation 21841584-672d-4710-bfc0-f5a8b1dbf116 · outbound

This paper cites - Each side of the square is equal to the height of the rectangle, which is given as 32 units.

From Answers to Rationales: Self-Aligning Multimodal Reasoning with Answer-Oriented Chain-of-Thought - Each side of the square is equal to the height of the rectangle, which is given as 32 units

Reference 74

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:19:45.078402Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-06T21:19:39.353941Z digest=sha256:d4c9050b047f73c756e91f2a6b171e6e15416ee764720208a6125990dc5fd85b

Observation fe9dc88b-a511-49e8-b5a3-7bb538349296 · outbound

This paper cites - The diagonal of a square with side length s is given by s\sqrt{2}.

From Answers to Rationales: Self-Aligning Multimodal Reasoning with Answer-Oriented Chain-of-Thought - The diagonal of a square with side length s is given by s\sqrt{2}

Reference 75

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:19:45.066148Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-06T21:19:39.555704Z digest=sha256:c7d26bb40f3170b90464e89743df58f34dd84da4321695b145dfaf750a19f010

Observation 75072758-ec69-4e69-ae37-13309c87ea6b · outbound

This paper cites objects": [ {.

From Answers to Rationales: Self-Aligning Multimodal Reasoning with Answer-Oriented Chain-of-Thought objects": [ {

Reference 76

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:19:45.055324Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-06T21:19:39.697546Z digest=sha256:ac57b512b84c443af39573ccfbc10877354f24d1c6e835e6d076cdb57d96b045

Observation ce7acc8a-7b65-4e0e-a640-cca607d7fb03 · outbound

This paper cites an unresolved cited work.

From Answers to Rationales: Self-Aligning Multimodal Reasoning with Answer-Oriented Chain-of-Thought Unresolved cited work

Reference 77

Resolution
unresolved
raw_fallback, observed 2026-08-06T21:19:45.045568Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-06T21:19:39.904367Z digest=sha256:38cd85855d843ef995d92dd2aedba32f2c0852d5eb6492a98d84a3c1a9760cd1

Observation 7b657692-e593-4303-be7e-ab141c48ca68 · outbound

This paper cites objects": [ {.

From Answers to Rationales: Self-Aligning Multimodal Reasoning with Answer-Oriented Chain-of-Thought objects": [ {

Reference 78

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:19:45.036316Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-06T21:19:40.066958Z digest=sha256:ba810563073974b68cd93979d2bd84705a9c91a9593c090c66a20288dcb70dd6

Observation ee271c14-4e87-4c91-bd70-311482409aea · outbound

This paper cites an unresolved cited work.

From Answers to Rationales: Self-Aligning Multimodal Reasoning with Answer-Oriented Chain-of-Thought Unresolved cited work

Reference 79

Resolution
unresolved
raw_fallback, observed 2026-08-06T21:19:45.026104Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-06T21:19:40.197225Z digest=sha256:0ed081a14a6c20d08d133c554acf655064e2443230dda153dd39d8aa8630ad3c

Observation f88a98fc-b701-4a4d-bfd9-f731d866947b · outbound

This paper cites an unresolved cited work.

From Answers to Rationales: Self-Aligning Multimodal Reasoning with Answer-Oriented Chain-of-Thought Unresolved cited work

Reference 80

Resolution
unresolved
raw_fallback, observed 2026-08-06T21:19:45.016416Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-06T21:19:40.353407Z digest=sha256:4e93931bb45c22c06af8a7d7dfbb9f241e44ed932c64ed7e52fece2b8969d055

Observation 1883a4f1-7c59-4490-bf2f-b2cf1a354f1c · outbound

This paper cites an unresolved cited work.

From Answers to Rationales: Self-Aligning Multimodal Reasoning with Answer-Oriented Chain-of-Thought Unresolved cited work

Reference 81

Resolution
unresolved
raw_fallback, observed 2026-08-06T21:19:45.007140Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-06T21:19:40.505961Z digest=sha256:d143d50245e2f946a65450ad06da2496dd0afe188c79e5d6ced5125872bcd088

Observation 2ad79d24-12a9-441e-91b7-affaaf63f698 · outbound

This paper cites an unresolved cited work.

From Answers to Rationales: Self-Aligning Multimodal Reasoning with Answer-Oriented Chain-of-Thought Unresolved cited work

Reference 82

Resolution
unresolved
raw_fallback, observed 2026-08-06T21:19:44.997450Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-06T21:19:40.674246Z digest=sha256:74859f2fe5fab1f67ed9e0deafd20f6d858f959e7602d40ae69d131aa8040788

Observation d0b55959-a5ab-490a-98d2-0e3beca719df · outbound

This paper cites an unresolved cited work.

From Answers to Rationales: Self-Aligning Multimodal Reasoning with Answer-Oriented Chain-of-Thought Unresolved cited work

Reference 83

Resolution
unresolved
raw_fallback, observed 2026-08-06T21:19:44.988969Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-06T21:19:40.913540Z digest=sha256:1e56457e427432395f26d0309a5499a81a0e87b3386f003664a010d066251e06

Observation c222e408-bde4-4718-89c1-13c4082e2752 · outbound

This paper cites an unresolved cited work.

From Answers to Rationales: Self-Aligning Multimodal Reasoning with Answer-Oriented Chain-of-Thought Unresolved cited work

Reference 84

Resolution
unresolved
raw_fallback, observed 2026-08-06T21:19:44.981093Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-06T21:19:41.053299Z digest=sha256:865469dedaabda3ecf5ad5f1aa10e0c5ba9437e25acabfd79eb18fbaffc892bf

Observation e7512077-b686-4bd8-a9b3-5cea7199c551 · outbound

This paper cites an unresolved cited work.

From Answers to Rationales: Self-Aligning Multimodal Reasoning with Answer-Oriented Chain-of-Thought Unresolved cited work

Reference 85

Resolution
unresolved
raw_fallback, observed 2026-08-06T21:19:44.972431Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-06T21:19:41.231965Z digest=sha256:99ad1ae7eae45d47dfb60880fcc26835189c7e0984e5d33791d2599368b63693

Observation 56bc3bf0-545e-492a-9f06-abd4b9e3ffb5 · outbound

This paper cites objects": [{.

From Answers to Rationales: Self-Aligning Multimodal Reasoning with Answer-Oriented Chain-of-Thought objects": [{

Reference 86

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:19:44.964871Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-06T21:19:41.411657Z digest=sha256:8bf645ce10225eb34a212cad24110dbdeef94c66fb0f0d4db23ae235214b04bd

Observation f684a584-5991-4a67-a1aa-b443994bcc49 · outbound

This paper cites **E** - Early blastocyst 3.

From Answers to Rationales: Self-Aligning Multimodal Reasoning with Answer-Oriented Chain-of-Thought **E** - Early blastocyst 3

Reference 87

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:19:44.957213Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-06T21:19:41.562213Z digest=sha256:e6ce4362fc28c8525d92d1811a182fb9b0dff66b84b022697a409dbfb085fa5f

Observation 11dbcc7f-07d4-4bfd-9fd0-89dcdfd5d89f · outbound

This paper cites (D) No.

From Answers to Rationales: Self-Aligning Multimodal Reasoning with Answer-Oriented Chain-of-Thought (D) No

Reference 88

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:19:44.948916Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-06T21:19:41.753763Z digest=sha256:cf65ea15fad1413dbed6e172f02bbd69510e9e16019b2949bd392b5b367aff22

Observation 7448bd87-3e1f-46a6-9bd6-fc8b96006fee · outbound

This paper cites an unresolved cited work.

From Answers to Rationales: Self-Aligning Multimodal Reasoning with Answer-Oriented Chain-of-Thought Unresolved cited work

Reference 89

Resolution
unresolved
raw_fallback, observed 2026-08-06T21:19:44.939965Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-06T21:19:41.888306Z digest=sha256:e341defbf7df4fc6a09f164e4e59613ed903d95df15d5702f262115978faf3e3

Observation fc7a62d8-836e-4b9c-85a4-3c8aa8f6a1cd · outbound

This paper cites an unresolved cited work.

From Answers to Rationales: Self-Aligning Multimodal Reasoning with Answer-Oriented Chain-of-Thought Unresolved cited work

Reference 90

Resolution
unresolved
raw_fallback, observed 2026-08-06T21:19:44.930051Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-06T21:19:42.048251Z digest=sha256:27e0962f260f6906c3b6ab0aa1fd34721a7680ea249735f8921822b87b56aa29

Observation e6b28067-6d4a-4129-a304-bffe378bbf2b · outbound

This paper cites an unresolved cited work.

From Answers to Rationales: Self-Aligning Multimodal Reasoning with Answer-Oriented Chain-of-Thought Unresolved cited work

Reference 91

Resolution
unresolved
raw_fallback, observed 2026-08-06T21:19:44.920005Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-06T21:19:42.250601Z digest=sha256:2163c1d4a8136cbd92703d41e646bf3964b97b623a5a2d7876e514e8e4e26176

Observation 52c84aee-7ee7-4961-a5b7-29a97ddc5936 · outbound

This paper cites an unresolved cited work.

From Answers to Rationales: Self-Aligning Multimodal Reasoning with Answer-Oriented Chain-of-Thought Unresolved cited work

Reference 92

Resolution
unresolved
raw_fallback, observed 2026-08-06T21:19:44.850249Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-06T21:19:42.419828Z digest=sha256:00735c279d2566c4cc433e4e83734ef3278aacd5d97f2fe52f04acfe389f909d

Observation ab215d8c-386d-424d-8cac-6d20dfabc0da · outbound

This paper cites an unresolved cited work.

From Answers to Rationales: Self-Aligning Multimodal Reasoning with Answer-Oriented Chain-of-Thought Unresolved cited work

Reference 93

Resolution
unresolved
raw_fallback, observed 2026-08-06T21:19:44.638857Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-06T21:19:42.629321Z digest=sha256:abad18bf738e20612411655807600ae4dcfde4de5089f7bac534eefb3bb05fa5

Observation bdeacc47-2d6a-4866-ae5d-8a665d175dc6 · outbound

This paper cites objects": [ {.

From Answers to Rationales: Self-Aligning Multimodal Reasoning with Answer-Oriented Chain-of-Thought objects": [ {

Reference 94

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:19:44.429933Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-06T21:19:42.743404Z digest=sha256:044e41afb15f14f32a58d0a19c200d0a7985565eee3e025801b64ce75f30849e

Observation 35053a44-ae4e-4854-9b81-98b7c152a411 · outbound

This paper cites an unresolved cited work.

From Answers to Rationales: Self-Aligning Multimodal Reasoning with Answer-Oriented Chain-of-Thought Unresolved cited work

Reference 95

Resolution
unresolved
raw_fallback, observed 2026-08-06T21:19:44.238427Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-06T21:19:42.783517Z digest=sha256:b411b64400f7271859e067b3651a32b4203e54081e92a64da63e450ab0c0e696

Observation 5b2c26af-e05c-445d-ad92-00dc4fc1c7ec · outbound

This paper cites an unresolved cited work.

From Answers to Rationales: Self-Aligning Multimodal Reasoning with Answer-Oriented Chain-of-Thought Unresolved cited work

Reference 96

Resolution
unresolved
raw_fallback, observed 2026-08-06T21:19:44.075489Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-06T21:19:42.836296Z digest=sha256:9c266e63a89d26748732faac1b6ada1cb712fced16e747cfdbb8fc4da4d2bbbf

Observation 0700ebef-9854-4246-8fec-40276333546d · outbound

This paper cites an unresolved cited work.

From Answers to Rationales: Self-Aligning Multimodal Reasoning with Answer-Oriented Chain-of-Thought Unresolved cited work

Reference 97

Resolution
unresolved
raw_fallback, observed 2026-08-06T21:19:43.777934Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-06T21:19:42.927117Z digest=sha256:ed8ea6cf5bd3f6c84082a65a8be6666bebf4b70575b013cb237559c04c4af20e

Observation 5da17c2f-9656-45aa-af6a-8b064383b604 · outbound

This paper cites The correct answer is: (C) smaller than Qwen-VL-7B + SMART : The derivative of the function y = log_2(x) is \frac{1}{x ln(2)}.

From Answers to Rationales: Self-Aligning Multimodal Reasoning with Answer-Oriented Chain-of-Thought The correct answer is: (C) smaller than Qwen-VL-7B + SMART : The derivative of the function y = log_2(x) is \frac{1}{x ln(2)}

Reference 98

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T21:19:43.630021Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-06T21:19:42.977063Z digest=sha256:6d1b6afb58ed21dd64cb8e71bbbd1e706beb046d4161f0f201f2c5d279b3985a

Pith citing papers

No inbound Pith citation observations are available.