Pith. sign in

Paper Citation Record · LEDGER

MPCC: A Novel Benchmark for Multimodal Planning with Complex Constraints in Multimodal Large Language Models

As of 8 August 2026, this Paper Citation Record lists 45 of 45 outbound references and 0 inbound Pith citation observations for arXiv:2507.23382.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2507.23382 v1

Coverage vector

measured 45 of 45 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-06T10:53:20.673277Z

measured 45 of 45 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

45 of 45 outbound references displayed

  • verified exact1
  • verified fuzzy24
  • unresolved20
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 38f6022c-3e77-4d45-b575-8e1ee440fa5a · outbound

This paper cites Gpt-4 technical report.

MPCC: A Novel Benchmark for Multimodal Planning with Complex Constraints in Multimodal Large Language Models Gpt-4 technical report

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:53:21.519968Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T10:53:20.497524Z digest=sha256:51101537047e73cb9ee842ae09aa220a94c43ced8ed61742ff87a5f738fafea2

Observation dfc13025-eb1c-4a87-88b3-7fc3ede17fa9 · outbound

This paper cites Towards Reasoning Era: A Survey of Long Chain-of-Thought for Reasoning Large Language Models.

MPCC: A Novel Benchmark for Multimodal Planning with Complex Constraints in Multimodal Large Language Models Towards Reasoning Era: A Survey of Long Chain-of-Thought for Reasoning Large Language Models

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-06T10:53:20.502152Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T10:53:20.502152Z digest=sha256:893f35799ba43f028d24750f9ef86c7165562ec92829a2b90f05decb9265d7a7

Observation 7d024f42-d35d-4b7d-92ee-16babd56faaf · outbound

This paper cites Ecm: A unified electronic circuit model for explaining the emergence of in-context learning and chain-of-thought in large language model.arXiv preprint arXiv:2502.03325, 2025.

MPCC: A Novel Benchmark for Multimodal Planning with Complex Constraints in Multimodal Large Language Models Ecm: A unified electronic circuit model for explaining the emergence of in-context learning and chain-of-thought in large language model.arXiv preprint arXiv:2502.03325, 2025

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-06T10:53:20.506901Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T10:53:20.506901Z digest=sha256:8a12886755174a8fc54cfa27458467a48e72528b5166881957f5a291ba9e4bc8

Observation 13f70436-3fe5-434d-90c7-e59b42e12ad9 · outbound

This paper cites Un- locking the capabilities of thought: A reasoning boundary framework to quantify and optimize chain-of-thought.

MPCC: A Novel Benchmark for Multimodal Planning with Complex Constraints in Multimodal Large Language Models Un- locking the capabilities of thought: A reasoning boundary framework to quantify and optimize chain-of-thought

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:53:21.507106Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T10:53:20.511551Z digest=sha256:fae24823b6415f377fa5a223e1408a09d4d55ad7da39996b47d3b31fd047f453

Observation 7d3e216e-b486-4320-a1fe-67023c648907 · outbound

This paper cites M 3 cot: A novel benchmark for multi-domain multi-step multi-modal chain-of- thought.

MPCC: A Novel Benchmark for Multimodal Planning with Complex Constraints in Multimodal Large Language Models M 3 cot: A novel benchmark for multi-domain multi-step multi-modal chain-of- thought

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:53:21.493455Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T10:53:20.515644Z digest=sha256:f804d8b1a3274ef2fc35e3523241c2863a7010520cb5cb4162c986499f202726

Observation 4a376083-56d5-406a-b71c-5c2c975fcac7 · outbound

This paper cites Egoplan-bench: Benchmarking egocentric embodied planning with multimodal large language models.CoRR, 2023.

MPCC: A Novel Benchmark for Multimodal Planning with Complex Constraints in Multimodal Large Language Models Egoplan-bench: Benchmarking egocentric embodied planning with multimodal large language models.CoRR, 2023

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:53:21.479982Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T10:53:20.519787Z digest=sha256:c806a3c5e4b37315f9c8887aa613f020fcd8d1dc2cbb11c2872e035b86d1308b

Observation 03ca3a9d-8cff-4683-9d69-15435d998f64 · outbound

This paper cites Internvl: Scaling up vision foundation models and aligning for generic visual-linguistic tasks.

MPCC: A Novel Benchmark for Multimodal Planning with Complex Constraints in Multimodal Large Language Models Internvl: Scaling up vision foundation models and aligning for generic visual-linguistic tasks

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:53:21.464734Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T10:53:20.524370Z digest=sha256:f5ee1f0da4e9915c9fb90163596915972285c68c2643c28903449855eb74a406

Observation 71c564ea-345e-4f30-a3cc-ae076c1446f7 · outbound

This paper cites Visual thoughts: A unified perspective of understanding multimodal chain-of-thought.arXiv preprint arXiv:2505.15510, 2025.

MPCC: A Novel Benchmark for Multimodal Planning with Complex Constraints in Multimodal Large Language Models Visual thoughts: A unified perspective of understanding multimodal chain-of-thought.arXiv preprint arXiv:2505.15510, 2025

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-06T10:53:20.528393Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T10:53:20.528393Z digest=sha256:5be618151a93d645478e99e93f3adb58a849c39de0275afdd309a3e87bb41cc5

Observation f1690525-f2cc-429f-a8e7-a80942b28811 · outbound

This paper cites Comt: A novel benchmark for chain of multi-modal thought on large vision-language models.

MPCC: A Novel Benchmark for Multimodal Planning with Complex Constraints in Multimodal Large Language Models Comt: A novel benchmark for chain of multi-modal thought on large vision-language models

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:53:21.451184Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T10:53:20.532223Z digest=sha256:8f3e03d2fbdbc26404d0d7a0f7b70303440f8b0daf972e4eab96e4d0faeac912

Observation c27c78a9-7cd2-4fe4-9320-b1a8af5c589a · outbound

This paper cites A Survey on In-context Learning.

MPCC: A Novel Benchmark for Multimodal Planning with Complex Constraints in Multimodal Large Language Models A Survey on In-context Learning

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-06T10:53:20.535987Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T10:53:20.535987Z digest=sha256:8b99ef87517ff01340172f69a727d1749992d0eb7dd34fd8f25bfef0db65e935

Observation 5d2493a2-9ab8-4f45-b2f6-a97da016bfb0 · outbound

This paper cites Vlmevalkit: An open- source toolkit for evaluating large multi-modality models.

MPCC: A Novel Benchmark for Multimodal Planning with Complex Constraints in Multimodal Large Language Models Vlmevalkit: An open- source toolkit for evaluating large multi-modality models

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:53:21.438119Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T10:53:20.540500Z digest=sha256:ddf8bffa1145a30c2909ca599d91d1ea08db327f84d18caf17ecb8b423df320a

Observation ef8c7bb5-9147-4a2e-8615-507acd7add9b · outbound

This paper cites DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning.

MPCC: A Novel Benchmark for Multimodal Planning with Complex Constraints in Multimodal Large Language Models DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-06T10:53:20.544678Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T10:53:20.544678Z digest=sha256:e050048793d296f2058000a7be40c4ee74ff1118bce8f771d3d2f728b7e99fe2

Observation e9cfd071-65c8-4457-a939-78342128c09f · outbound

This paper cites Mllm-compbench: A compara- tive reasoning benchmark for multimodal llms.

MPCC: A Novel Benchmark for Multimodal Planning with Complex Constraints in Multimodal Large Language Models Mllm-compbench: A compara- tive reasoning benchmark for multimodal llms

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:53:21.423840Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T10:53:20.548904Z digest=sha256:334cd89806eae779ea4170b27e45f09ffd57c693730c978767890b24ac8a94fb

Observation e13ced5d-1622-4e0f-827e-a450a91147f7 · outbound

This paper cites VisualWebArena: Evaluating Multimodal Agents on Realistic Visual Web Tasks.

MPCC: A Novel Benchmark for Multimodal Planning with Complex Constraints in Multimodal Large Language Models VisualWebArena: Evaluating Multimodal Agents on Realistic Visual Web Tasks

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-06T10:53:20.552721Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T10:53:20.552721Z digest=sha256:7860a5f3c04676800e223761513e70eaf82088ee68b8b2528ce02fcdcae48423

Observation 88aabee3-7b5a-49bd-a6d3-f42ca4afb14d · outbound

This paper cites Tree search for language model agents, 2024.

MPCC: A Novel Benchmark for Multimodal Planning with Complex Constraints in Multimodal Large Language Models Tree search for language model agents, 2024

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:53:21.410865Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T10:53:20.556992Z digest=sha256:42f6adbe45bc4527f0bff4bface84d18fa1657a59c04827e6d93ffbef2c5e563

Observation b3647e32-7be7-471a-a0b0-f4b92f9b4980 · outbound

This paper cites Large language models are zero-shot reasoners, 2022.

MPCC: A Novel Benchmark for Multimodal Planning with Complex Constraints in Multimodal Large Language Models Large language models are zero-shot reasoners, 2022

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-06T10:53:20.560926Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T10:53:20.560926Z digest=sha256:9508934d84c8beee5a6c23083789a036983c2487dcdb24814a8375283cc459c3

Observation 07d3ce67-45a7-4c52-8f9e-b251c317ace6 · outbound

This paper cites Llava-onevision: Easy visual task transfer, 2024.

MPCC: A Novel Benchmark for Multimodal Planning with Complex Constraints in Multimodal Large Language Models Llava-onevision: Easy visual task transfer, 2024

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-06T10:53:20.564766Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T10:53:20.564766Z digest=sha256:6e55cece3317d3057a90855e8562142944dae2e5233b48f1e51e9f00cd97eddb

Observation f00ebf18-11f2-4a6c-95a1-5e526fe0e5a2 · outbound

This paper cites Seed-bench: Benchmarking multimodal large language models.

MPCC: A Novel Benchmark for Multimodal Planning with Complex Constraints in Multimodal Large Language Models Seed-bench: Benchmarking multimodal large language models

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:53:21.379683Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T10:53:20.568470Z digest=sha256:0a27f1951247c30d0e70b4743a946063e69d5773d0522194ab7856f9f44472fb

Observation f2526b6e-39a8-41ed-b38b-492360bdee11 · outbound

This paper cites Behavior-1k: A benchmark for embodied ai with 1,000 everyday activities and realistic simulation.

MPCC: A Novel Benchmark for Multimodal Planning with Complex Constraints in Multimodal Large Language Models Behavior-1k: A benchmark for embodied ai with 1,000 everyday activities and realistic simulation

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-06T10:53:20.572100Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T10:53:20.572100Z digest=sha256:cded9255575d8ffe1b9f93dfcaf6590d67b8fcab4b17148e6b3f4de673828dac

Observation 20607dfc-846e-443e-8f8b-8454ead93101 · outbound

This paper cites Ferret-ui 2: Mastering universal user interface understanding across platforms, 2024.

MPCC: A Novel Benchmark for Multimodal Planning with Complex Constraints in Multimodal Large Language Models Ferret-ui 2: Mastering universal user interface understanding across platforms, 2024

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:53:21.357032Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T10:53:20.576092Z digest=sha256:6e6669e2d6a5bed852d0b9596aba8e9a137f6b0cb026f00b35f58429f3f5257d

Observation 80832cb4-1839-4aab-a06d-2149fd8c6fc6 · outbound

This paper cites RoboMP$^2$: A Robotic Multimodal Perception-Planning Framework with Multimodal Large Language Models.

MPCC: A Novel Benchmark for Multimodal Planning with Complex Constraints in Multimodal Large Language Models RoboMP$^2$: A Robotic Multimodal Perception-Planning Framework with Multimodal Large Language Models

Reference 21

Resolution
verified exact
local_arxiv, observed 2026-08-06T10:53:20.857484Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T10:53:20.579657Z digest=sha256:a78d82b10419ee4de23db8138a6824caf6dd80daa5a8eb7b308c8a100cd4f031

Observation b2cb7bce-9da2-403c-a023-6a9b43a115c2 · outbound

This paper cites JanusFlow: Harmonizing Autoregression and Rectified Flow for Unified Multimodal Understanding and Generation.

MPCC: A Novel Benchmark for Multimodal Planning with Complex Constraints in Multimodal Large Language Models JanusFlow: Harmonizing Autoregression and Rectified Flow for Unified Multimodal Understanding and Generation

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-06T10:53:20.583635Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T10:53:20.583635Z digest=sha256:ce9dca159e789355d7f80412c82c341da80ad90a05f2550af22aaff72fa13918

Observation 2d7c84d5-0e23-4ef3-90af-b17f03632481 · outbound

This paper cites m & m’s: A benchmark to evaluate tool-use for m ulti-step m ulti-modal tasks.

MPCC: A Novel Benchmark for Multimodal Planning with Complex Constraints in Multimodal Large Language Models m & m’s: A benchmark to evaluate tool-use for m ulti-step m ulti-modal tasks

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:53:21.342611Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T10:53:20.587553Z digest=sha256:404e242b00c43d95d4f728502711b3f7a865178711453f9191f6e3bc8a8de2ec

Observation e18d92fa-a9c9-4f16-9239-96f48aad6669 · outbound

This paper cites Perception test: A diagnostic benchmark for multimodal video models.

MPCC: A Novel Benchmark for Multimodal Planning with Complex Constraints in Multimodal Large Language Models Perception test: A diagnostic benchmark for multimodal video models

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:53:21.329216Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T10:53:20.591223Z digest=sha256:04d3667e4764d25d3928fb44b3c049ef7d860051882d9f2c2d6b2bc0a8aec7be

Observation 3b11e0bc-ef44-4a5c-b56b-fb42a421b34a · outbound

This paper cites What Factors Affect Multi-Modal In-Context Learning? An In-Depth Exploration.

MPCC: A Novel Benchmark for Multimodal Planning with Complex Constraints in Multimodal Large Language Models What Factors Affect Multi-Modal In-Context Learning? An In-Depth Exploration

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-06T10:53:20.594935Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T10:53:20.594935Z digest=sha256:fd807876b532a5d69caaea8069013e8101fad3514ca14f5d221598f3ce1d4950

Observation e1663673-4322-4823-b7ba-ef621e7f5a66 · outbound

This paper cites Mementos: System support for long-running computation on rfid-scale devices.

MPCC: A Novel Benchmark for Multimodal Planning with Complex Constraints in Multimodal Large Language Models Mementos: System support for long-running computation on rfid-scale devices

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:53:21.315978Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T10:53:20.598656Z digest=sha256:a668d703df571337f25fcab27d10ec0f635c60e2a7221c280feb0f4529754009

Observation fdb76c80-db26-41c4-824f-56969aaf932b · outbound

This paper cites Alfred: A benchmark for interpreting grounded instructions for everyday tasks.

MPCC: A Novel Benchmark for Multimodal Planning with Complex Constraints in Multimodal Large Language Models Alfred: A benchmark for interpreting grounded instructions for everyday tasks

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:53:21.302254Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T10:53:20.602259Z digest=sha256:b7a850787ebd59eb5952dd854bcd7d34ab552a4f5a2f815f4dd46318368c6f17

Observation 4300e1d5-9a33-4489-a7f5-b3b29cf42ef6 · outbound

This paper cites Gemini: A Family of Highly Capable Multimodal Models.

MPCC: A Novel Benchmark for Multimodal Planning with Complex Constraints in Multimodal Large Language Models Gemini: A Family of Highly Capable Multimodal Models

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-06T10:53:20.606796Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T10:53:20.606796Z digest=sha256:8cc05d4cf1a4588d6e941eeb9c707b0c3437b281100fe72702473763cc34a611

Observation d6e7aed5-07cf-4a4c-988a-d32486c2c349 · outbound

This paper cites Qvq: To see the world with wisdom, 2024.

MPCC: A Novel Benchmark for Multimodal Planning with Complex Constraints in Multimodal Large Language Models Qvq: To see the world with wisdom, 2024

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:53:21.287495Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T10:53:20.610677Z digest=sha256:1763023899d9c817e66a11a3e2d4162b566651832174f78d466e006c01228850

Observation 415ebad2-08aa-44e8-8832-4fc3bfd011b0 · outbound

This paper cites Planbench: An extensible benchmark for evaluating large language models on planning and reasoning about change, 2023.

MPCC: A Novel Benchmark for Multimodal Planning with Complex Constraints in Multimodal Large Language Models Planbench: An extensible benchmark for evaluating large language models on planning and reasoning about change, 2023

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-06T10:53:20.614626Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T10:53:20.614626Z digest=sha256:5da46acd296ee3cb94aecf16710e3657e9cf7ab0610d2d765420e68dddfc9979

Observation 175ce598-3c26-41fd-a595-99bdebe0ef8d · outbound

This paper cites Xlrs-bench: Could your multimodal llms understand extremely large ultra-high-resolution remote sensing imagery?, 2025.

MPCC: A Novel Benchmark for Multimodal Planning with Complex Constraints in Multimodal Large Language Models Xlrs-bench: Could your multimodal llms understand extremely large ultra-high-resolution remote sensing imagery?, 2025

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:53:21.264906Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T10:53:20.618401Z digest=sha256:0ef8487494d83a440912252a3b4d2f1c7af4ac8360f061bb09064f96e5dcbbc8

Observation 64d6f275-87ee-4f1c-944e-b2e104f60ad3 · outbound

This paper cites Plan-and-Solve Prompting: Improving Zero-Shot Chain-of-Thought Reasoning by Large Language Models.

MPCC: A Novel Benchmark for Multimodal Planning with Complex Constraints in Multimodal Large Language Models Plan-and-Solve Prompting: Improving Zero-Shot Chain-of-Thought Reasoning by Large Language Models

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-06T10:53:20.622273Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T10:53:20.622273Z digest=sha256:59c0f5d9d135de09138fb1720065040d592294060b656773dc0e20e7d15bf1f4

Observation a43f15a6-d6dd-4aec-9a8f-94076181df54 · outbound

This paper cites Qwen2-VL: Enhancing Vision-Language Model's Perception of the World at Any Resolution.

MPCC: A Novel Benchmark for Multimodal Planning with Complex Constraints in Multimodal Large Language Models Qwen2-VL: Enhancing Vision-Language Model's Perception of the World at Any Resolution

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-06T10:53:20.626682Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T10:53:20.626682Z digest=sha256:32f8e23af885174daf8d870941d53c64ec6c8281ab2a8c703c47e6752a2cc94f

Observation e25a1a96-ca37-46a9-a969-d45c7963374c · outbound

This paper cites S3 agent: Unlocking the power of vllm for zero-shot multi-modal sarcasm detection.ACM Transactions on Multimedia Computing, Communications and Applications, 2024.

MPCC: A Novel Benchmark for Multimodal Planning with Complex Constraints in Multimodal Large Language Models S3 agent: Unlocking the power of vllm for zero-shot multi-modal sarcasm detection.ACM Transactions on Multimedia Computing, Communications and Applications, 2024

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:53:21.251436Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T10:53:20.630864Z digest=sha256:db4c4d8c35b386939b30fa337ec89b7b219aee4dacb1fbcc18dafcac737653de

Observation 31d4ee66-0bea-495b-95af-e3346d405260 · outbound

This paper cites Chain-of-thought prompting elicits reasoning in large language models.

MPCC: A Novel Benchmark for Multimodal Planning with Complex Constraints in Multimodal Large Language Models Chain-of-thought prompting elicits reasoning in large language models

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:53:21.237662Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T10:53:20.634494Z digest=sha256:bf463ec74fcb789a6c2bbe76535a3e5f5ab960b804da6dae97206a102b72b8ae

Observation e098dcdc-292b-4deb-9cea-d21e07d50b92 · outbound

This paper cites Janus: Decoupling Visual Encoding for Unified Multimodal Understanding and Generation.

MPCC: A Novel Benchmark for Multimodal Planning with Complex Constraints in Multimodal Large Language Models Janus: Decoupling Visual Encoding for Unified Multimodal Understanding and Generation

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-06T10:53:20.638489Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T10:53:20.638489Z digest=sha256:1218ee7bee5038e8fa1095c5f772bb350efb396ff9a3d84c6892fb950d52aed5

Observation 8dad09c2-d8f5-4d26-a6bf-c430b5b79682 · outbound

This paper cites DeepSeek-VL2: Mixture-of-Experts Vision-Language Models for Advanced Multimodal Understanding.

MPCC: A Novel Benchmark for Multimodal Planning with Complex Constraints in Multimodal Large Language Models DeepSeek-VL2: Mixture-of-Experts Vision-Language Models for Advanced Multimodal Understanding

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-06T10:53:20.642627Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T10:53:20.642627Z digest=sha256:3a8cfcb5715ecedc2901b1af172de06044a2299e2eb82c0f56222383829092e0

Observation 40164780-a8fd-448a-b20b-6e7c10104746 · outbound

This paper cites LogicVista: Multimodal LLM Logical Reasoning Benchmark in Visual Contexts.

MPCC: A Novel Benchmark for Multimodal Planning with Complex Constraints in Multimodal Large Language Models LogicVista: Multimodal LLM Logical Reasoning Benchmark in Visual Contexts

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-06T10:53:20.646437Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T10:53:20.646437Z digest=sha256:17dd3a6765a1f9839ea53d9dbb07f919d6dc069bb084c20ca8cef43143f9d98f

Observation 9fbe474f-7f41-4e34-8c20-c37f16660a2e · outbound

This paper cites Osworld: Benchmarking multimodal agents for open-ended tasks in real computer envi- ronments.

MPCC: A Novel Benchmark for Multimodal Planning with Complex Constraints in Multimodal Large Language Models Osworld: Benchmarking multimodal agents for open-ended tasks in real computer envi- ronments

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:53:21.222379Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T10:53:20.650321Z digest=sha256:795a631c65415066627c00aa9d0a5754a64c3aae32ee2c151ffb2d1addbbf70a

Observation 85312a74-d55f-422f-afab-0aded02c8247 · outbound

This paper cites Mm-react: Prompt- ing chatgpt for multimodal reasoning and action, 2023.

MPCC: A Novel Benchmark for Multimodal Planning with Complex Constraints in Multimodal Large Language Models Mm-react: Prompt- ing chatgpt for multimodal reasoning and action, 2023

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:53:21.208589Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T10:53:20.653970Z digest=sha256:ace62b0c3294cb287faeee335e1a81a99e19dda551a6fb255ac91737aa662fdd

Observation 26a3a3dc-d9ba-4060-ba7c-8bd3e211a70c · outbound

This paper cites MMT-Bench: A Comprehensive Multimodal Benchmark for Evaluating Large Vision-Language Models Towards Multitask AGI.

MPCC: A Novel Benchmark for Multimodal Planning with Complex Constraints in Multimodal Large Language Models MMT-Bench: A Comprehensive Multimodal Benchmark for Evaluating Large Vision-Language Models Towards Multitask AGI

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-06T10:53:20.657760Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T10:53:20.657760Z digest=sha256:fa62d28b582c164542b7d9bb0239ff1446fc793d772591587272259bc6f803a3

Observation bb719f0c-e043-4363-9b71-a54dc0b829d1 · outbound

This paper cites Mmmu: A massive multi-discipline multimodal understanding and reasoning benchmark for expert agi.

MPCC: A Novel Benchmark for Multimodal Planning with Complex Constraints in Multimodal Large Language Models Mmmu: A massive multi-discipline multimodal understanding and reasoning benchmark for expert agi

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:53:21.194795Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T10:53:20.661802Z digest=sha256:b8dd8d4993f7c5a5f545c85c23f6b63803c9601bd504dc6bed2d68613cf45dd3

Observation df3b235e-10dd-4c5d-a6ab-30a1188db3d8 · outbound

This paper cites Open3dvqa: A benchmark for comprehensive spatial reasoning with multimodal large language model in open space, 2025.

MPCC: A Novel Benchmark for Multimodal Planning with Complex Constraints in Multimodal Large Language Models Open3dvqa: A benchmark for comprehensive spatial reasoning with multimodal large language model in open space, 2025

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:53:21.180213Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T10:53:20.665367Z digest=sha256:e4ad5a8297b818c17fda7ca95b4c05bce7ca0c0ba9bb689370f145e611ac0225

Observation 54c936e4-ef99-4a58-a577-45b9aa36009a · outbound

This paper cites Le, Ed H.

MPCC: A Novel Benchmark for Multimodal Planning with Complex Constraints in Multimodal Large Language Models Le, Ed H

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:53:21.167018Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T10:53:20.669515Z digest=sha256:f6d98ec195fcc245dfaff5668524d321c475aac77df817ed4bbb6cbf5bc41f48

Observation bb16c00f-dbbe-4255-910c-68e64810b32f · outbound

This paper cites WebArena: A Realistic Web Environment for Building Autonomous Agents.

MPCC: A Novel Benchmark for Multimodal Planning with Complex Constraints in Multimodal Large Language Models WebArena: A Realistic Web Environment for Building Autonomous Agents

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-06T10:53:20.673277Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T10:53:20.673277Z digest=sha256:d445204e0f137473993321cf996511f1e3fa08508663eba35815715b3e8f679d

Pith citing papers

No inbound Pith citation observations are available.