Pith. sign in

Paper Citation Record · LEDGER

Mixture-of-Visual-Thoughts: Exploring Context-Adaptive Reasoning Mode Selection for General Visual Reasoning

As of 6 August 2026, this Paper Citation Record lists 65 of 65 outbound references and 0 inbound Pith citation observations for arXiv:2509.22746.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2509.22746 v2

Coverage vector

measured 65 of 65 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-05-18T13:33:12.508639Z

measured 65 of 65 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-06T06:34:29.942622+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

65 of 65 outbound references displayed

  • verified exact38
  • verified fuzzy23
  • unresolved1
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch3

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation eae4980e-6bef-4e83-8138-d448e7efe1f3 · outbound

This paper cites write newline.

Mixture-of-Visual-Thoughts: Exploring Context-Adaptive Reasoning Mode Selection for General Visual Reasoning write newline

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T13:36:26.750252Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-05-18T13:33:12.508639Z digest=sha256:5c8b65c344fc8038d348bce27fec36e0aa6c33b8c9c336b70d7284c5f8e78189

Observation c8ac538d-0065-4f3b-a0db-1d6e6af1e94c · outbound

This paper cites Qwen2.5-VL Technical Report.

Mixture-of-Visual-Thoughts: Exploring Context-Adaptive Reasoning Mode Selection for General Visual Reasoning Qwen2.5-VL Technical Report

Reference 2

Resolution
verified exact
local_arxiv, observed 2026-05-18T13:36:25.232148Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-05-18T13:33:12.508639Z digest=sha256:f9cc8eabf4ffcdc48d4556ae3e577cb1a8ff273783a21c6c6dc9da8747a43d63

Observation 17c759d6-33ce-4c0c-81a8-af07493f16c6 · outbound

This paper cites Graph of thoughts: Solving elaborate problems with large language models.

Mixture-of-Visual-Thoughts: Exploring Context-Adaptive Reasoning Mode Selection for General Visual Reasoning Graph of thoughts: Solving elaborate problems with large language models

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T13:36:26.743189Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-05-18T13:33:12.508639Z digest=sha256:0f14dcdbe0a993e1ad0184d9437c2c0b44ca5a71a9b437264fd487114f1e6c63

Observation 6d1a59b4-9003-4387-9eaa-8ccc8389a82b · outbound

This paper cites Ground- r1: Incentivizing grounded visual reasoning via reinforcement learning.arXiv preprint arXiv:2505.20272.

Mixture-of-Visual-Thoughts: Exploring Context-Adaptive Reasoning Mode Selection for General Visual Reasoning Ground- r1: Incentivizing grounded visual reasoning via reinforcement learning.arXiv preprint arXiv:2505.20272

Reference 4

Resolution
verified exact
arxiv_id, observed 2026-05-18T13:36:25.149784Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-05-18T13:33:12.508639Z digest=sha256:b02715f39c91175f79b138e22b05b54ab37cdf52068b419d5d93b6415881ebba

Observation d4b59fc0-9878-4ddd-876b-91f6ccbf08e0 · outbound

This paper cites SFT or RL? An Early Investigation into Training R1-Like Reasoning Large Vision-Language Models.

Mixture-of-Visual-Thoughts: Exploring Context-Adaptive Reasoning Mode Selection for General Visual Reasoning SFT or RL? An Early Investigation into Training R1-Like Reasoning Large Vision-Language Models

Reference 5

Resolution
verified exact
local_arxiv, observed 2026-05-18T13:36:25.155140Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-05-18T13:33:12.508639Z digest=sha256:6a71d161159ee56a38fb469afd533a3713bfd11ed931a0c19141ec3adf6ac7be

Observation f194457a-7a47-4c76-b4ce-a86eb070ed54 · outbound

This paper cites Shikra: Unleashing Multimodal LLM's Referential Dialogue Magic.

Mixture-of-Visual-Thoughts: Exploring Context-Adaptive Reasoning Mode Selection for General Visual Reasoning Shikra: Unleashing Multimodal LLM's Referential Dialogue Magic

Reference 6

Resolution
verified exact
local_arxiv, observed 2026-05-18T13:36:25.302737Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-05-18T13:33:12.508639Z digest=sha256:189af0400aa51b17ce9fb264f702d67ff31f1071348b2fd2cd041c638ac30043

Observation 6ef69faa-0bc5-496b-8918-bf34aa4904a2 · outbound

This paper cites Are we on the right way for evaluating large vision-language models? Advances in Neural Information Processing Systems, 37: 0 27056--27087, 2024 a.

Mixture-of-Visual-Thoughts: Exploring Context-Adaptive Reasoning Mode Selection for General Visual Reasoning Are we on the right way for evaluating large vision-language models? Advances in Neural Information Processing Systems, 37: 0 27056--27087, 2024 a

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T13:36:26.739617Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-05-18T13:33:12.508639Z digest=sha256:0e0e8b4464dfeb44a80dbd2e3a654ac1460b2fee8903e29c4ed35bae640e95f8

Observation e43bead3-ca1c-4d9d-9253-4a27eda87d19 · outbound

This paper cites How far are we to gpt-4v? closing the gap to commercial multimodal models with open-source suites.

Mixture-of-Visual-Thoughts: Exploring Context-Adaptive Reasoning Mode Selection for General Visual Reasoning How far are we to gpt-4v? closing the gap to commercial multimodal models with open-source suites

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T13:42:39.682307Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-05-18T13:33:12.508639Z digest=sha256:30d6897825b481316b099fe3647a0ea7aed22c2ead41189441195d6c5ae245af

Observation 3e86c49b-ab93-4d7c-8b77-484c4265308e · outbound

This paper cites DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning.

Mixture-of-Visual-Thoughts: Exploring Context-Adaptive Reasoning Mode Selection for General Visual Reasoning DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning

Reference 9

Resolution
verified exact
local_arxiv, observed 2026-05-18T13:36:25.236412Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-05-18T13:33:12.508639Z digest=sha256:c4f05ac11595326432e965eb764505ed7fb8792cee2106bbb8b4fb2531f2b61e

Observation 8bbbb5f5-c80c-480a-8c5d-981d46d98f93 · outbound

This paper cites Insight-v: Exploring long-chain visual reasoning with multimodal large language models.

Mixture-of-Visual-Thoughts: Exploring Context-Adaptive Reasoning Mode Selection for General Visual Reasoning Insight-v: Exploring long-chain visual reasoning with multimodal large language models

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T13:36:26.746641Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-05-18T13:33:12.508639Z digest=sha256:5f674d238cdce4353130a840b6b1d0587ff464688b0365e74116a9b9d36f3ac1

Observation d21b6467-f76a-4848-8bce-95109d40f558 · outbound

This paper cites Virgo: A Preliminary Exploration on Reproducing o1-like MLLM.

Mixture-of-Visual-Thoughts: Exploring Context-Adaptive Reasoning Mode Selection for General Visual Reasoning Virgo: A Preliminary Exploration on Reproducing o1-like MLLM

Reference 11

Resolution
verified exact
arxiv_id, observed 2026-05-18T13:36:25.184297Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-05-18T13:33:12.508639Z digest=sha256:17790fdcbe29ae6211473d8236e8cf183b906cf08331b051234ca6317ffff0a5

Observation 9c9dc810-3210-4559-aafd-a0dafc2aa787 · outbound

This paper cites GRIT: Teaching MLLMs to Think with Images.

Mixture-of-Visual-Thoughts: Exploring Context-Adaptive Reasoning Mode Selection for General Visual Reasoning GRIT: Teaching MLLMs to Think with Images

Reference 12

Resolution
verified exact
local_arxiv, observed 2026-05-18T13:36:25.109996Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-05-18T13:33:12.508639Z digest=sha256:393cac3f13a0443a3d50ddd3bc979b571d7bc45d772b42511be3340befe0c1aa

Observation 3f8f19ae-d2cf-40d4-ab00-d60404a94583 · outbound

This paper cites G-llava: Solving geometric problem with multi-modal large language model.

Mixture-of-Visual-Thoughts: Exploring Context-Adaptive Reasoning Mode Selection for General Visual Reasoning G-llava: Solving geometric problem with multi-modal large language model

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T13:36:26.730466Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-05-18T13:33:12.508639Z digest=sha256:24ed4e6f9cc718403326a0bd6d1d0497a44c1759c97394f3684fa350c162a46a

Observation 0aed07cb-5226-4026-b352-fb2322e0c3bb · outbound

This paper cites Vision-R1: Incentivizing Reasoning Capability in Multimodal Large Language Models.

Mixture-of-Visual-Thoughts: Exploring Context-Adaptive Reasoning Mode Selection for General Visual Reasoning Vision-R1: Incentivizing Reasoning Capability in Multimodal Large Language Models

Reference 14

Resolution
verified exact
local_arxiv, observed 2026-05-18T13:36:25.213344Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-05-18T13:33:12.508639Z digest=sha256:324f44f3a831675969d047275792d0f8b8c819bfcc9f49e36df309c214a38468

Observation 7fa6fbd7-27a6-4211-ae90-4efd5d9c972e · outbound

This paper cites GPT-4o System Card.

Mixture-of-Visual-Thoughts: Exploring Context-Adaptive Reasoning Mode Selection for General Visual Reasoning GPT-4o System Card

Reference 15

Resolution
verified exact
local_arxiv, observed 2026-05-18T13:36:25.105758Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-05-18T13:33:12.508639Z digest=sha256:3ebd5dd493827989289a4af7c34ef1b17b1936aa390ac81573f3b0641fafcb22

Observation e46a6aa9-4bfd-4f67-82dd-d0f155e4dc36 · outbound

This paper cites Kimi k1.5: Scaling Reinforcement Learning with LLMs.

Mixture-of-Visual-Thoughts: Exploring Context-Adaptive Reasoning Mode Selection for General Visual Reasoning Kimi k1.5: Scaling Reinforcement Learning with LLMs

Reference 16

Resolution
verified exact
local_arxiv, observed 2026-05-18T13:36:25.170186Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-05-18T13:33:12.508639Z digest=sha256:66fcb3622d13b7b73a800285ce672cb644575d95b7e5fae30e3154e3b6da0b58

Observation 8973cd9b-1d6a-4b9c-b762-eb876efbe796 · outbound

This paper cites Large language models are zero-shot reasoners.

Mixture-of-Visual-Thoughts: Exploring Context-Adaptive Reasoning Mode Selection for General Visual Reasoning Large language models are zero-shot reasoners

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T13:36:26.723765Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-05-18T13:33:12.508639Z digest=sha256:a00616aab384b3482873be42c391e345a06969cc15b0346f6df441abcbfcc3c4

Observation bb16ecd9-18b3-442d-aff4-02166106f159 · outbound

This paper cites Hypertree proof search for neural theorem proving.

Mixture-of-Visual-Thoughts: Exploring Context-Adaptive Reasoning Mode Selection for General Visual Reasoning Hypertree proof search for neural theorem proving

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T13:36:26.727128Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-05-18T13:33:12.508639Z digest=sha256:fc97467e347166fdd338f0c2637af8289aaf375674ceb1cc196ee2b5bd646b4b

Observation 596e9e5d-faa1-4a5c-8635-1f1fdec22f2d · outbound

This paper cites Scaffolding coordinates to promote vision-language coordination in large multi-modal models.

Mixture-of-Visual-Thoughts: Exploring Context-Adaptive Reasoning Mode Selection for General Visual Reasoning Scaffolding coordinates to promote vision-language coordination in large multi-modal models

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T13:36:26.736572Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-05-18T13:33:12.508639Z digest=sha256:1774e3be4f4f6fc9196290d3816beb427bc05ebf14a92239ea0c6b58421b9804

Observation e21b9d30-ad94-4c8a-89a7-2b52dcf8ce34 · outbound

This paper cites LLaVA-OneVision: Easy Visual Task Transfer.

Mixture-of-Visual-Thoughts: Exploring Context-Adaptive Reasoning Mode Selection for General Visual Reasoning LLaVA-OneVision: Easy Visual Task Transfer

Reference 20

Resolution
verified exact
local_arxiv, observed 2026-05-18T13:36:25.193876Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-05-18T13:33:12.508639Z digest=sha256:c853fee65d95805ebd7b023b1f5e40622193ea14a73067406c5f26a32f1ff205

Observation 6dd36938-4398-4374-b00a-c05e0a2b47a0 · outbound

This paper cites Numinamath.

Mixture-of-Visual-Thoughts: Exploring Context-Adaptive Reasoning Mode Selection for General Visual Reasoning Numinamath

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T13:36:26.733370Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-05-18T13:33:12.508639Z digest=sha256:9cda3012e69771cd288cc8d70a349af15f5de3c7c0850433e222c39b8254aa72

Observation 6d9d66cb-4387-4d7b-ba27-34cfb4bc259c · outbound

This paper cites Evaluating Object Hallucination in Large Vision-Language Models.

Mixture-of-Visual-Thoughts: Exploring Context-Adaptive Reasoning Mode Selection for General Visual Reasoning Evaluating Object Hallucination in Large Vision-Language Models

Reference 22

Resolution
verified exact
local_arxiv, observed 2026-05-18T13:36:25.174850Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-05-18T13:33:12.508639Z digest=sha256:772d83b66346344f2ef868078527dfa331c0552570e27212a11d67762d9993c7

Observation 494333cc-0eec-4838-ba19-b8d7b4637278 · outbound

This paper cites VoCoT: Unleashing Visually Grounded Multi-Step Reasoning in Large Multi-Modal Models.

Mixture-of-Visual-Thoughts: Exploring Context-Adaptive Reasoning Mode Selection for General Visual Reasoning VoCoT: Unleashing Visually Grounded Multi-Step Reasoning in Large Multi-Modal Models

Reference 23

Resolution
verified exact
arxiv_id, observed 2026-05-18T13:36:25.203478Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-05-18T13:33:12.508639Z digest=sha256:420f349fb9c11d226765fcf2b4fb18ff1f774b87c94889c7038ba4cba9737ec0

Observation 55ab3a8a-ffba-4052-b05a-8e2206cdbf30 · outbound

This paper cites Visual-RFT: Visual Reinforcement Fine-Tuning.

Mixture-of-Visual-Thoughts: Exploring Context-Adaptive Reasoning Mode Selection for General Visual Reasoning Visual-RFT: Visual Reinforcement Fine-Tuning

Reference 24

Resolution
verified exact
local_arxiv, observed 2026-05-18T13:36:25.298265Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-05-18T13:33:12.508639Z digest=sha256:feb310890e55276f04c15aaa075d6b9c4fc91337f589b1b2010fa5f6034e65d9

Observation 86a720e7-bcee-4441-a8da-013578a46e74 · outbound

This paper cites MathVista: Evaluating Mathematical Reasoning of Foundation Models in Visual Contexts.

Mixture-of-Visual-Thoughts: Exploring Context-Adaptive Reasoning Mode Selection for General Visual Reasoning MathVista: Evaluating Mathematical Reasoning of Foundation Models in Visual Contexts

Reference 25

Resolution
verified exact
local_arxiv, observed 2026-05-18T13:36:25.143460Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-05-18T13:33:12.508639Z digest=sha256:26dd72ab1f4002f0861ca0401d27a9baaaf7e7868d36b829f88dda2a12a2e0c9

Observation 489fd87b-cca6-4c56-a5fd-ae0c8460da40 · outbound

This paper cites One RL to See Them All: Visual Triple Unified Reinforcement Learning.

Mixture-of-Visual-Thoughts: Exploring Context-Adaptive Reasoning Mode Selection for General Visual Reasoning One RL to See Them All: Visual Triple Unified Reinforcement Learning

Reference 26

Resolution
verified exact
local_arxiv, observed 2026-05-18T13:36:25.286858Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-05-18T13:33:12.508639Z digest=sha256:89e943cee3c7eff9aa5f3b9275c395960de413fc2c9d18e7edbf799e0d7457d3

Observation eadc6229-3e3b-4bc8-923f-6196c9449831 · outbound

This paper cites Self-refine: Iterative refinement with self-feedback.

Mixture-of-Visual-Thoughts: Exploring Context-Adaptive Reasoning Mode Selection for General Visual Reasoning Self-refine: Iterative refinement with self-feedback

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T13:42:39.671092Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-05-18T13:33:12.508639Z digest=sha256:f3222ed32dd77dfaf33fda93fcbd2d1721ba05d65c06384d266ff2e74c6c46d0

Observation 9f1c250c-57b9-4e58-8e2c-4a66ec98a420 · outbound

This paper cites MM-Eureka: Exploring the Frontiers of Multimodal Reasoning with Rule-based Reinforcement Learning.

Mixture-of-Visual-Thoughts: Exploring Context-Adaptive Reasoning Mode Selection for General Visual Reasoning MM-Eureka: Exploring the Frontiers of Multimodal Reasoning with Rule-based Reinforcement Learning

Reference 28

Resolution
verified exact
local_arxiv, observed 2026-05-18T13:36:25.130294Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-05-18T13:33:12.508639Z digest=sha256:7ff0d53e9d98987e3957f3e48d8e4ec8b5eac3d9466702993d9a3fb635fed82d

Observation 613234ac-408c-4fae-919e-eb474ca6ee8f · outbound

This paper cites Compositional chain-of-thought prompting for large multimodal models.

Mixture-of-Visual-Thoughts: Exploring Context-Adaptive Reasoning Mode Selection for General Visual Reasoning Compositional chain-of-thought prompting for large multimodal models

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T13:42:39.663341Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-05-18T13:33:12.508639Z digest=sha256:f91b05a33fa75d97665f4180ccbacb0c2d7bcd7c5b3d9ead116f1c7fd7cc4a60

Observation f9f65693-3908-4838-91d7-c64f85198f84 · outbound

This paper cites Omnicount: Multi-label object counting with semantic-geometric priors.

Mixture-of-Visual-Thoughts: Exploring Context-Adaptive Reasoning Mode Selection for General Visual Reasoning Omnicount: Multi-label object counting with semantic-geometric priors

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T13:42:39.656261Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-05-18T13:33:12.508639Z digest=sha256:da8e89ad729bd69098b4d2472632bce516881d174db1eb3690bfdc85dedbcaa1

Observation 671b24a3-b707-41dc-bda2-00e89fc29b94 · outbound

This paper cites Thinking with images.

Mixture-of-Visual-Thoughts: Exploring Context-Adaptive Reasoning Mode Selection for General Visual Reasoning Thinking with images

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T13:42:39.667756Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-05-18T13:33:12.508639Z digest=sha256:48f8f285a4e4485a7d022874d5fb45369a6aa7feee4cc3dbdb9cd562d20975c9

Observation 54bc6d70-db0e-4b97-9749-2feecf785e46 · outbound

This paper cites LMM-R1: Empowering 3B LMMs with Strong Reasoning Abilities Through Two-Stage Rule-Based RL.

Mixture-of-Visual-Thoughts: Exploring Context-Adaptive Reasoning Mode Selection for General Visual Reasoning LMM-R1: Empowering 3B LMMs with Strong Reasoning Abilities Through Two-Stage Rule-Based RL

Reference 32

Resolution
verified exact
local_arxiv, observed 2026-05-18T13:36:25.114501Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-05-18T13:33:12.508639Z digest=sha256:0342d982e63ee254e38867656c872442cb88d82f380f024f16e7794fabb11134

Observation c6303bcc-b9d2-4b01-a448-de45481f3088 · outbound

This paper cites We-Math: Does Your Large Multimodal Model Achieve Human-like Mathematical Reasoning?.

Mixture-of-Visual-Thoughts: Exploring Context-Adaptive Reasoning Mode Selection for General Visual Reasoning We-Math: Does Your Large Multimodal Model Achieve Human-like Mathematical Reasoning?

Reference 33

Resolution
verified exact
local_arxiv, observed 2026-05-18T13:36:25.124560Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-05-18T13:33:12.508639Z digest=sha256:b65688800e50358857b08c7560fad6d1cf2a59705997fdbb0deb354bdd721bb5

Observation 0e8882d0-2c75-484b-93d9-0e90df5dc9bd · outbound

This paper cites QwQ-32B : Embracing the power of reinforcement learning, March 2025.

Mixture-of-Visual-Thoughts: Exploring Context-Adaptive Reasoning Mode Selection for General Visual Reasoning QwQ-32B : Embracing the power of reinforcement learning, March 2025

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T13:42:39.674868Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-05-18T13:33:12.508639Z digest=sha256:146527ae2097505dbb352d3470381f8fd575d9994fa0ba2ac4af0e37bf4e0073

Observation eb758a0e-2145-4ff9-afca-dceab286a150 · outbound

This paper cites Grounded Reinforcement Learning for Visual Reasoning.

Mixture-of-Visual-Thoughts: Exploring Context-Adaptive Reasoning Mode Selection for General Visual Reasoning Grounded Reinforcement Learning for Visual Reasoning

Reference 35

Resolution
verified exact
arxiv_id, observed 2026-05-18T13:36:25.189444Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-05-18T13:33:12.508639Z digest=sha256:464a36021ebcd401e5d9bab93a54c404d99d5493cbbf65de19d160ced75022c2

Observation 2aa5a4b8-039d-4a8d-936d-578fd348553f · outbound

This paper cites Visual CoT: Advancing Multi-Modal Language Models with a Comprehensive Dataset and Benchmark for Chain-of-Thought Reasoning.

Mixture-of-Visual-Thoughts: Exploring Context-Adaptive Reasoning Mode Selection for General Visual Reasoning Visual CoT: Advancing Multi-Modal Language Models with a Comprehensive Dataset and Benchmark for Chain-of-Thought Reasoning

Reference 36

Resolution
verified exact
arxiv_id, observed 2026-05-18T13:36:25.165701Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-05-18T13:33:12.508639Z digest=sha256:95f318749cfc3d58b6563789a232803027e11eeb31504d3c8dcf53212e8463b6

Observation d55c139d-dd28-4bd1-965c-94d5930f9d24 · outbound

This paper cites DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models.

Mixture-of-Visual-Thoughts: Exploring Context-Adaptive Reasoning Mode Selection for General Visual Reasoning DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models

Reference 37

Resolution
verified exact
local_arxiv, observed 2026-05-18T13:36:25.198881Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-05-18T13:33:12.508639Z digest=sha256:7ce3b42155391ff62fdec10d841f05881352415a58467c5fc15f777d2bd6d681

Observation 6b4bfc76-4bee-4a75-ac32-64f54fd4d0b2 · outbound

This paper cites ZoomEye: Enhancing Multimodal LLMs with Human-Like Zooming Capabilities through Tree-Based Image Exploration.

Mixture-of-Visual-Thoughts: Exploring Context-Adaptive Reasoning Mode Selection for General Visual Reasoning ZoomEye: Enhancing Multimodal LLMs with Human-Like Zooming Capabilities through Tree-Based Image Exploration

Reference 38

Resolution
verified exact
arxiv_id, observed 2026-05-18T13:36:25.241161Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-05-18T13:33:12.508639Z digest=sha256:24b2f082a2794c04c5a631c78c2191a6850f5e25d54c09fef117d3b21340ed69

Observation df381330-8b9c-4ad3-9593-fbc176df5458 · outbound

This paper cites Reflexion: Language agents with verbal reinforcement learning.

Mixture-of-Visual-Thoughts: Exploring Context-Adaptive Reasoning Mode Selection for General Visual Reasoning Reflexion: Language agents with verbal reinforcement learning

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T13:42:39.678876Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-05-18T13:33:12.508639Z digest=sha256:44cf4f02806ca87aaa1a2ff8ecd2dd3a163bedf6b5d6b219c2ecf8ae8b90a0ef

Observation 848dac7b-ef8a-4f42-a83e-ab8cb60f7440 · outbound

This paper cites LlamaV-o1: Rethinking Step-by-step Visual Reasoning in LLMs.

Mixture-of-Visual-Thoughts: Exploring Context-Adaptive Reasoning Mode Selection for General Visual Reasoning LlamaV-o1: Rethinking Step-by-step Visual Reasoning in LLMs

Reference 40

Resolution
metadata mismatch
arxiv_id, observed 2026-05-18T13:36:25.179472Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-05-18T13:33:12.508639Z digest=sha256:0e4988e8f4248652b2cece50598777016c3a3bc9f7ed31b1acb64031d572616b

Observation 34c7873a-49d0-4992-aa6d-4e9c1cc6747c · outbound

This paper cites Toward self-improvement of llms via imagination, searching, and criticizing.

Mixture-of-Visual-Thoughts: Exploring Context-Adaptive Reasoning Mode Selection for General Visual Reasoning Toward self-improvement of llms via imagination, searching, and criticizing

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T13:36:26.772917Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-05-18T13:33:12.508639Z digest=sha256:579efe90366ba08bfeab9947af5b3676acc62af70987546ddade4f0858a89173

Observation 96f4f656-afad-4a21-94d7-8d4589147942 · outbound

This paper cites Measuring Multimodal Mathematical Reasoning with MATH-Vision Dataset.

Mixture-of-Visual-Thoughts: Exploring Context-Adaptive Reasoning Mode Selection for General Visual Reasoning Measuring Multimodal Mathematical Reasoning with MATH-Vision Dataset

Reference 42

Resolution
verified exact
local_arxiv, observed 2026-05-18T13:36:25.223356Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-05-18T13:33:12.508639Z digest=sha256:9c8ef94d8079b9376e0a565b80693c72558e6fc900177db0e4ae32c63dca1b2d

Observation 58d32fa8-b084-48b7-a943-96adcba97d2a · outbound

This paper cites Self-consistency improves chain of thought reasoning in language models.

Mixture-of-Visual-Thoughts: Exploring Context-Adaptive Reasoning Mode Selection for General Visual Reasoning Self-consistency improves chain of thought reasoning in language models

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T13:36:26.764897Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-05-18T13:33:12.508639Z digest=sha256:8188373edbbfd2965fac5ea8e6ea1152b9255e2b4dd8571afc1120199d6a18e7

Observation 567823d3-0618-4525-af76-ca165c7c2c7a · outbound

This paper cites Chain-of-thought prompting elicits reasoning in large language models.

Mixture-of-Visual-Thoughts: Exploring Context-Adaptive Reasoning Mode Selection for General Visual Reasoning Chain-of-thought prompting elicits reasoning in large language models

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T13:36:26.769023Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-05-18T13:33:12.508639Z digest=sha256:d13b534f12f78ead5f62b0009a4b4ac7bcb02206ac49ea98212899843dafb5e4

Observation 0706187b-201c-4eee-8fa5-6dc413ea7157 · outbound

This paper cites Open vision reasoner: Transferring linguistic cognitive behavior for visual reasoning.

Mixture-of-Visual-Thoughts: Exploring Context-Adaptive Reasoning Mode Selection for General Visual Reasoning Open vision reasoner: Transferring linguistic cognitive behavior for visual reasoning

Reference 45

Resolution
verified exact
arxiv_id, observed 2026-05-18T13:36:25.136959Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-05-18T13:33:12.508639Z digest=sha256:ce5a07444912133437439ebe1275200e4bbf6ec41833cacb2ccc430b136e06a3

Observation f669f1a8-8023-4c2a-8d36-def6a0b6de81 · outbound

This paper cites SpatialScore: Towards Comprehensive Evaluation for Spatial Intelligence.

Mixture-of-Visual-Thoughts: Exploring Context-Adaptive Reasoning Mode Selection for General Visual Reasoning SpatialScore: Towards Comprehensive Evaluation for Spatial Intelligence

Reference 47

Resolution
verified exact
local_arxiv, observed 2026-05-18T13:36:25.264926Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-05-18T13:33:12.508639Z digest=sha256:b2efa508fd1b12ada4872f657a1332d8299075181d4c373a6869284729d3761b

Observation e64e0d55-fe91-403e-8dc5-24610c1afd0b · outbound

This paper cites V*: Guided Visual Search as a Core Mechanism in Multimodal LLMs.

Mixture-of-Visual-Thoughts: Exploring Context-Adaptive Reasoning Mode Selection for General Visual Reasoning V*: Guided Visual Search as a Core Mechanism in Multimodal LLMs

Reference 48

Resolution
verified exact
arxiv_id, observed 2026-05-18T13:36:25.246023Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-05-18T13:33:12.508639Z digest=sha256:001232aabd667d2cccdd01ad743b4b4d2a3b17f3e6e772775642d438fb75c513

Observation 17b940df-2077-4fa0-ac82-5b632675a6f1 · outbound

This paper cites Grounded Chain-of-Thought for Multimodal Large Language Models.

Mixture-of-Visual-Thoughts: Exploring Context-Adaptive Reasoning Mode Selection for General Visual Reasoning Grounded Chain-of-Thought for Multimodal Large Language Models

Reference 49

Resolution
verified exact
arxiv_id, observed 2026-05-18T13:36:25.208637Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-05-18T13:33:12.508639Z digest=sha256:510b107fcb7c7f49bd0d09ec3fde806bfa33980bfa56bf274288d0c82df52277

Observation 09915245-a471-48c9-899f-0195c2f18fa8 · outbound

This paper cites Self-evaluation guided beam search for reasoning.

Mixture-of-Visual-Thoughts: Exploring Context-Adaptive Reasoning Mode Selection for General Visual Reasoning Self-evaluation guided beam search for reasoning

Reference 50

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T13:36:26.776685Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-05-18T13:33:12.508639Z digest=sha256:075c18e3da85d86c71d74ae115f6c2a6a3cc4f01549f604df897d9df5e67d18d

Observation ef47cf23-1767-48a0-8c37-d3d5bfeceadf · outbound

This paper cites LLaVA-CoT: Let Vision Language Models Reason Step-by-Step.

Mixture-of-Visual-Thoughts: Exploring Context-Adaptive Reasoning Mode Selection for General Visual Reasoning LLaVA-CoT: Let Vision Language Models Reason Step-by-Step

Reference 51

Resolution
verified exact
local_arxiv, observed 2026-05-18T13:36:25.280777Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-05-18T13:33:12.508639Z digest=sha256:b91a5000093eae30cfcdefebfae7e1e5619c2e00ee904f21fcbb6c0e5248a7f8

Observation c2ff1905-bcb4-4ad8-b597-6d20402c7b01 · outbound

This paper cites GeoSense: Evaluating Identification and Application of Geometric Principles in Multimodal Reasoning.

Mixture-of-Visual-Thoughts: Exploring Context-Adaptive Reasoning Mode Selection for General Visual Reasoning GeoSense: Evaluating Identification and Application of Geometric Principles in Multimodal Reasoning

Reference 52

Resolution
verified exact
arxiv_id, observed 2026-05-18T13:36:25.255208Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-05-18T13:33:12.508639Z digest=sha256:acb0d952cd11fdd01af43d58d05be5bcd8e4ef033a773313c2b6c5016ee53770

Observation d0084653-492f-4f57-86f7-7dbd5c26c31e · outbound

This paper cites Set-of-mark prompting unleashes extraordinary visual grounding in gpt-4v.

Mixture-of-Visual-Thoughts: Exploring Context-Adaptive Reasoning Mode Selection for General Visual Reasoning Set-of-mark prompting unleashes extraordinary visual grounding in gpt-4v

Reference 53

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T13:36:26.761087Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-05-18T13:33:12.508639Z digest=sha256:fe39f27cbc038fd15891812ed0fafd9d18ff07d4aee532b26df01c6698f9e648

Observation ada1b670-4925-4ab9-8335-318ae0fea32d · outbound

This paper cites R1-Onevision: Advancing Generalized Multimodal Reasoning through Cross-Modal Formalization.

Mixture-of-Visual-Thoughts: Exploring Context-Adaptive Reasoning Mode Selection for General Visual Reasoning R1-Onevision: Advancing Generalized Multimodal Reasoning through Cross-Modal Formalization

Reference 54

Resolution
verified exact
local_arxiv, observed 2026-05-18T13:36:25.260535Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-05-18T13:33:12.508639Z digest=sha256:425eaee01617c31b0c2a922194dafb146738605f64a437f1e441ef91227a6195

Observation 1901e20f-0587-4281-904a-fc08c4dbeab0 · outbound

This paper cites Mulberry: Empowering MLLM with o1-like Reasoning and Reflection via Collective Monte Carlo Tree Search.

Mixture-of-Visual-Thoughts: Exploring Context-Adaptive Reasoning Mode Selection for General Visual Reasoning Mulberry: Empowering MLLM with o1-like Reasoning and Reflection via Collective Monte Carlo Tree Search

Reference 55

Resolution
verified exact
arxiv_id, observed 2026-05-18T13:36:25.218707Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-05-18T13:33:12.508639Z digest=sha256:430a99f80a519f7895f0d9747534ce238d7c7751d30b75caf6a3603e723cdfd2

Observation b1cb55ff-7440-413f-84dc-0803f795c833 · outbound

This paper cites Tree of thoughts: Deliberate problem solving with large language models.

Mixture-of-Visual-Thoughts: Exploring Context-Adaptive Reasoning Mode Selection for General Visual Reasoning Tree of thoughts: Deliberate problem solving with large language models

Reference 56

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T13:36:26.757761Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-05-18T13:33:12.508639Z digest=sha256:3bd7531107704dcf5dcf692ca0bda5a6be1a53e8eeb1656c666e4a9093ee4e31

Observation c9be5a1e-b53a-4b62-b4db-e4ca1eede66e · outbound

This paper cites DAPO: An Open-Source LLM Reinforcement Learning System at Scale.

Mixture-of-Visual-Thoughts: Exploring Context-Adaptive Reasoning Mode Selection for General Visual Reasoning DAPO: An Open-Source LLM Reinforcement Learning System at Scale

Reference 57

Resolution
verified exact
local_arxiv, observed 2026-05-18T13:36:25.307894Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-05-18T13:33:12.508639Z digest=sha256:071f11d3ed25eb0d4cf754559b83e454bdb68663fbea1c0ef208405c0af46e42

Observation 64f8fb67-8516-4565-a997-430f917769b6 · outbound

This paper cites MathVerse: Does Your Multi-modal LLM Truly See the Diagrams in Visual Math Problems?.

Mixture-of-Visual-Thoughts: Exploring Context-Adaptive Reasoning Mode Selection for General Visual Reasoning MathVerse: Does Your Multi-modal LLM Truly See the Diagrams in Visual Math Problems?

Reference 58

Resolution
verified exact
local_arxiv, observed 2026-05-18T13:36:25.269615Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-05-18T13:33:12.508639Z digest=sha256:f1c8a35895371694e75bede4a3064579e3743c7b2045ee07b2d27afe56b7f416

Observation 1c72cbf4-188f-4bfd-9167-af8f49785e6a · outbound

This paper cites Improve Vision Language Model Chain-of-thought Reasoning.

Mixture-of-Visual-Thoughts: Exploring Context-Adaptive Reasoning Mode Selection for General Visual Reasoning Improve Vision Language Model Chain-of-thought Reasoning

Reference 59

Resolution
metadata mismatch
arxiv_id, observed 2026-05-18T13:36:25.275632Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-05-18T13:33:12.508639Z digest=sha256:4a74ef44985927822278367de6f5f6d38647b1464cdec0ef2076c1fa085a0e43

Observation b35e14bb-6c22-4050-877b-00556a8a467d · outbound

This paper cites Adaptive Chain-of-Focus Reasoning via Dynamic Visual Search and Zooming for Efficient VLMs.

Mixture-of-Visual-Thoughts: Exploring Context-Adaptive Reasoning Mode Selection for General Visual Reasoning Adaptive Chain-of-Focus Reasoning via Dynamic Visual Search and Zooming for Efficient VLMs

Reference 60

Resolution
verified exact
local_arxiv, observed 2026-05-18T13:36:25.119251Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-05-18T13:33:12.508639Z digest=sha256:63fde3b096e97043eb49d47850abad3c55a6e1a90fd7c249bc9041e69f72d039

Observation e14bc0e6-da42-4368-9d66-5d4696bbf492 · outbound

This paper cites DeepEyes: Incentivizing "Thinking with Images" via Reinforcement Learning.

Mixture-of-Visual-Thoughts: Exploring Context-Adaptive Reasoning Mode Selection for General Visual Reasoning DeepEyes: Incentivizing "Thinking with Images" via Reinforcement Learning

Reference 61

Resolution
verified exact
local_arxiv, observed 2026-05-18T13:36:25.292168Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-05-18T13:33:12.508639Z digest=sha256:b206d10c1b236784046152f654b0eea3d4946746eff4cb720c844cb4b44e8507

Observation 0cf59d23-0df2-4584-a084-056771c44046 · outbound

This paper cites Image-of-Thought Prompting for Visual Reasoning Refinement in Multimodal Large Language Models.

Mixture-of-Visual-Thoughts: Exploring Context-Adaptive Reasoning Mode Selection for General Visual Reasoning Image-of-Thought Prompting for Visual Reasoning Refinement in Multimodal Large Language Models

Reference 62

Resolution
verified exact
arxiv_id, observed 2026-05-18T13:36:25.228065Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-05-18T13:33:12.508639Z digest=sha256:ea4bb9a6397d1183ae7ff18ac4533be3128805619126da80416b97862386ed40

Observation f714d8db-b659-4c1e-b422-80eb5910e66a · outbound

This paper cites InternVL3: Exploring Advanced Training and Test-Time Recipes for Open-Source Multimodal Models.

Mixture-of-Visual-Thoughts: Exploring Context-Adaptive Reasoning Mode Selection for General Visual Reasoning InternVL3: Exploring Advanced Training and Test-Time Recipes for Open-Source Multimodal Models

Reference 63

Resolution
verified exact
local_arxiv, observed 2026-05-18T13:36:25.250615Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-05-18T13:33:12.508639Z digest=sha256:2a499daefadfb28f23b5f5e16ac23937d8711b44c880ee24141f2e27950c736a

Observation fe667d02-bb25-4f8e-b3e8-a8bef73ff84b · outbound

This paper cites @esa (Ref.

Mixture-of-Visual-Thoughts: Exploring Context-Adaptive Reasoning Mode Selection for General Visual Reasoning @esa (Ref

Reference 64

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T13:42:39.652566Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-05-18T13:33:12.508639Z digest=sha256:f1fe075fc63c2d3e3840524b33cd96aab8e38e30fb6bd14c9a01a4f4abdb9702

Observation 3acfbb19-209a-46af-ba66-2e891a7e98ee · outbound

This paper cites an unresolved cited work.

Mixture-of-Visual-Thoughts: Exploring Context-Adaptive Reasoning Mode Selection for General Visual Reasoning Unresolved cited work

Reference 65

Resolution
unresolved
raw_fallback, observed 2026-05-18T13:36:26.754015Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-05-18T13:33:12.508639Z digest=sha256:9730484fd95d61f6778025ff9a9327826b08c7212b66e53dfa319ff69796e5a6

Observation dddb4ee2-b4b7-4bc8-b4bf-9f7c464dc0d2 · outbound

This paper cites (QGT+  o/߸ ;fQ Zt鐒gvZxG*J Y ȮY! dZs (HE E 2 n=#R.

Mixture-of-Visual-Thoughts: Exploring Context-Adaptive Reasoning Mode Selection for General Visual Reasoning (QGT+  o/߸ ;fQ Zt鐒gvZxG*J Y ȮY! dZs (HE E 2 n=#R

Reference 66

Resolution
metadata mismatch
arxiv_id, observed 2026-05-18T13:36:25.160912Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-05-18T13:33:12.508639Z digest=sha256:492178ef88a1f75974424484dae3f0fc126896421e7c85795fe9b8d467aefaeb

Pith citing papers

No inbound Pith citation observations are available.