Pith. sign in

Paper Citation Record · LEDGER

G-LLaVA: Solving Geometric Problem with Multi-Modal Large Language Model

As of 7 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 37 inbound Pith citation observations for arXiv:2312.11370.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2312.11370 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 37 of 37 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-06T06:34:29.942622+00:00

measured 37 of 37 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-06T20:25:23.198589Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-07-10T01:46:40.938849Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 6f71f00e-88eb-415c-b7e9-d9adf6a44215 · inbound

DeepSeek-VL: Towards Real-World Vision-Language Understanding cites this paper.

DeepSeek-VL: Towards Real-World Vision-Language Understanding G-LLaVA: Solving Geometric Problem with Multi-Modal Large Language Model

Reference 10

Resolution
verified exact
arxiv_id, observed 2026-05-11T17:58:54.443445Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-11T17:58:54.177359Z digest=sha256:69703b586bd0024be2da9aa3f61944c3493807b337d576148b00098163bb8d4f

Observation a2ede1a1-ee09-40b9-a87c-a1969a9c7436 · inbound

MathVerse: Does Your Multi-modal LLM Truly See the Diagrams in Visual Math Problems? cites this paper.

MathVerse: Does Your Multi-modal LLM Truly See the Diagrams in Visual Math Problems? G-LLaVA: Solving Geometric Problem with Multi-Modal Large Language Model

Reference 19

Resolution
metadata mismatch
arxiv_id, observed 2026-05-17T01:29:30.092967Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-17T01:29:30.032408Z digest=sha256:8a4f934bb0647f3e0da761f761bfdf6a2432c90257d98a974855fd0461ec213e

Observation babf1d04-1396-46e3-ad41-bf45371b67ff · inbound

Cambrian-1: A Fully Open, Vision-Centric Exploration of Multimodal LLMs cites this paper.

Cambrian-1: A Fully Open, Vision-Centric Exploration of Multimodal LLMs G-LLaVA: Solving Geometric Problem with Multi-Modal Large Language Model

Reference 44

Resolution
verified exact
arxiv_id, observed 2026-05-17T00:05:03.735798Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-17T00:05:03.547664Z digest=sha256:89908e7313c70bf25e6a5c4407992c20abc8bab5507b97d9f6f85639e92599ec

Observation 5ebba77d-4155-4dac-bc11-d8c1540cfc71 · inbound

We-Math: Does Your Large Multimodal Model Achieve Human-like Mathematical Reasoning? cites this paper.

We-Math: Does Your Large Multimodal Model Achieve Human-like Mathematical Reasoning? G-LLaVA: Solving Geometric Problem with Multi-Modal Large Language Model

Reference 29

Resolution
verified exact
arxiv_id, observed 2026-05-16T21:55:40.928589Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-16T21:55:40.808698Z digest=sha256:c5420488cdf866bd40011ffbe266e575b655fcec72a9682779702d7afc2d271b

Observation 4949e172-7d11-4456-b104-8a87ba7f15eb · inbound

CogVLM2: Visual Language Models for Image and Video Understanding cites this paper.

CogVLM2: Visual Language Models for Image and Video Understanding G-LLaVA: Solving Geometric Problem with Multi-Modal Large Language Model

Reference 16

Resolution
verified exact
arxiv_id, observed 2026-05-16T20:10:27.800589Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-16T20:10:27.633010Z digest=sha256:0b15524d562ebf279dfea89f883444d4a21acca877812cb212d0411a88d7685b

Observation e82f7f68-d45d-4bcc-80b1-51605fc0c417 · inbound

Enhancing the Reasoning Ability of Multimodal Large Language Models via Mixed Preference Optimization cites this paper.

Enhancing the Reasoning Ability of Multimodal Large Language Models via Mixed Preference Optimization G-LLaVA: Solving Geometric Problem with Multi-Modal Large Language Model

Reference 28

Resolution
verified exact
arxiv_id, observed 2026-05-16T09:16:17.352825Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-16T09:16:17.150383Z digest=sha256:edc1b8a9994d443b2b31fb60ba74bc36b79876642849eb07f7229dfb9fe41890

Observation 3c7de051-46ce-4805-8784-f20d08675b94 · inbound

NVILA: Efficient Frontier Visual Language Models cites this paper.

NVILA: Efficient Frontier Visual Language Models G-LLaVA: Solving Geometric Problem with Multi-Modal Large Language Model

Reference 121

Resolution
verified exact
arxiv_id, observed 2026-05-23T07:42:43.021604Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-23T07:42:22.478647Z digest=sha256:eb89b21c1b13cebff07ec8022bef8304584475aee737979e391e899607aeb968

Observation 0947025f-7df0-4489-8b06-76abe7dbd372 · inbound

MetaMorph: Multimodal Understanding and Generation via Instruction Tuning cites this paper.

MetaMorph: Multimodal Understanding and Generation via Instruction Tuning G-LLaVA: Solving Geometric Problem with Multi-Modal Large Language Model

Reference 123

Resolution
metadata mismatch
arxiv_id, observed 2026-05-17T07:51:13.363069Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-05-17T07:51:12.953777Z digest=sha256:bdb788a82c0a6785b1611918b6e1ba3aaf4aab5d57d7c336052195b987ca4241

Observation 253e581e-4786-46fb-aa88-3ec223495e9e · inbound

MathFlow: Enhancing the Perceptual Flow of MLLMs for Visual Mathematical Problems cites this paper.

MathFlow: Enhancing the Perceptual Flow of MLLMs for Visual Mathematical Problems G-LLaVA: Solving Geometric Problem with Multi-Modal Large Language Model

Reference 21

Resolution
verified exact
arxiv_id, observed 2026-05-22T22:57:13.389593Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-22T22:55:34.238427Z digest=sha256:3c1b625d10385a06254dd2ac2ac2642a038ccea21b4081c07804b06a863b18f6

Observation 61dc8cef-711d-445e-a750-d26e01ba6868 · inbound

FLARE: Fully Integration of Vision-Language Representations for Deep Cross-Modal Understanding cites this paper.

FLARE: Fully Integration of Vision-Language Representations for Deep Cross-Modal Understanding G-LLaVA: Solving Geometric Problem with Multi-Modal Large Language Model

Reference 20

Resolution
verified exact
arxiv_id, observed 2026-05-22T19:52:01.957115Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-22T19:49:00.961388Z digest=sha256:7a9f669c0e756ec059589ef8eda9feba75c8c582b7e76a12b18e413333802e01

Observation 99c40049-2da8-48ca-908d-407fabca8222 · inbound

InternVL3: Exploring Advanced Training and Test-Time Recipes for Open-Source Multimodal Models cites this paper.

InternVL3: Exploring Advanced Training and Test-Time Recipes for Open-Source Multimodal Models G-LLaVA: Solving Geometric Problem with Multi-Modal Large Language Model

Reference 40

Resolution
verified exact
arxiv_id, observed 2026-05-10T13:41:08.303127Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-10T13:41:07.991012Z digest=sha256:015fb7f85964073ef462b40d1de5444400dbd3598d5777428680df426a335e3c

Observation b4254f2d-c8c8-47ac-8d55-b62cdf06a913 · inbound

Vision-EKIPL: External Knowledge-Infused Policy Learning for Visual Reasoning cites this paper.

Vision-EKIPL: External Knowledge-Infused Policy Learning for Visual Reasoning G-LLaVA: Solving Geometric Problem with Multi-Modal Large Language Model

Reference 12

Resolution
verified exact
arxiv_id, observed 2026-05-19T10:37:15.031032Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-19T10:34:48.849524Z digest=sha256:75bc9772c057af58cd1ee71e915ac59bb5de7b22f560487a8e37529e78a5bd87

Observation c6576118-44e3-4e17-b879-850e3cd6a662 · inbound

Multimodal Mathematical Reasoning with Diverse Solving Perspective cites this paper.

Multimodal Mathematical Reasoning with Diverse Solving Perspective G-LLaVA: Solving Geometric Problem with Multi-Modal Large Language Model

Reference 59

Resolution
unresolved
no resolver link, observed 2026-08-06T20:25:23.198589Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:25:23.198589Z digest=sha256:8b0e2c602908b60a47e15b2d4982e2e8f1463e063af4f9f7e6c3ab307d067017

Observation 4a0d7780-f503-478c-aa5f-f25049a25352 · inbound

Correspondence as Video: Test-Time Adaption on SAM2 for Reference Segmentation in the Wild cites this paper.

Correspondence as Video: Test-Time Adaption on SAM2 for Reference Segmentation in the Wild G-LLaVA: Solving Geometric Problem with Multi-Modal Large Language Model

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-05T21:53:38.437566Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T21:53:38.437566Z digest=sha256:695b0267e6179a4cfec767d368cad1351ce2eccefd830241c84f54c0bd6556bf

Observation 9260a6bc-1ea0-4302-ab9e-0ad5f21dac99 · inbound

LaRe: Latent Refocusing for Multimodal Reasoning cites this paper.

LaRe: Latent Refocusing for Multimodal Reasoning G-LLaVA: Solving Geometric Problem with Multi-Modal Large Language Model

Reference 93

Resolution
unresolved
no resolver link, observed 2026-08-04T00:15:21.671366Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T00:15:21.671366Z digest=sha256:cbff7640ff1ba8e41e645a06bd72defaf5e8df843edbeaf1587ba88f6ec8415b

Observation 196a9300-bdf5-49ae-8fde-90af45385b61 · inbound

Toward an Artificial General Teacher: Procedural Geometry Data Generation and Visual Grounding with Vision-Language Models cites this paper.

Toward an Artificial General Teacher: Procedural Geometry Data Generation and Visual Grounding with Vision-Language Models G-LLaVA: Solving Geometric Problem with Multi-Modal Large Language Model

Reference 6

Resolution
verified exact
arxiv_id, observed 2026-05-13T20:38:14.635341Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-05-13T20:35:42.626519Z digest=sha256:6ef41ce129b02d20b3c718351b45cf6e3e689f0a76f68e774964f1699692c1a2

Observation e6fadf7e-0798-4dc7-a537-4da3a351b485 · inbound

E-VLA: Event-Augmented Vision-Language-Action Model for Dark and Blurred Scenes cites this paper.

E-VLA: Event-Augmented Vision-Language-Action Model for Dark and Blurred Scenes G-LLaVA: Solving Geometric Problem with Multi-Modal Large Language Model

Reference 11

Resolution
unresolved
no resolver link, observed 2026-07-13T09:40:35.631188Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-13T09:40:35.631188Z digest=sha256:1b59e0a18dcd4c9574589b31e5d1fdd9b9cd918321a2394e32377d7cac8068bd

Observation 7662a5ee-f7d9-4156-9e8e-b20e19bc200b · inbound

Less Detail, Better Answers: Degradation-Driven Prompting for VQA cites this paper.

Less Detail, Better Answers: Degradation-Driven Prompting for VQA G-LLaVA: Solving Geometric Problem with Multi-Modal Large Language Model

Reference 11

Resolution
verified exact
arxiv_id, observed 2026-05-10T22:05:47.913119Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-10T20:17:01.867903Z digest=sha256:7e2a91427b1e01babe748827cad4515c4fc3c98caa5ad95b8d4d924346fa34ed

Observation 061d0f40-b6e0-4d7b-b126-b513cfc8e4b2 · inbound

DT2IT-MRM: Debiased Preference Construction and Iterative Training for Multimodal Reward Modeling cites this paper.

DT2IT-MRM: Debiased Preference Construction and Iterative Training for Multimodal Reward Modeling G-LLaVA: Solving Geometric Problem with Multi-Modal Large Language Model

Reference 11

Resolution
verified exact
arxiv_id, observed 2026-05-11T13:26:03.821164Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-10T01:45:30.001398Z digest=sha256:6046a3ddaeef201f86b58fd9d54eb9213917ff2fd7f2d8df0907d3f1d6100a12

Observation 8356b553-0e6b-4723-b791-f5b587dd76cb · inbound

ZAYA1-VL-8B Technical Report cites this paper.

ZAYA1-VL-8B Technical Report G-LLaVA: Solving Geometric Problem with Multi-Modal Large Language Model

Reference 169

Resolution
verified exact
arxiv_id, observed 2026-05-12T08:21:23.398223Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-12T01:15:16.607346Z digest=sha256:ae5361dcf78388524c0a058781dc261bb32d0243192df1a2f8e3f2f36df5a48c

Observation d87bcc52-cdef-44ef-a5f8-4b25c959765b · inbound

Are Tools Always Beneficial? Learning to Invoke Tools Adaptively for Dual-Mode Multimodal LLM Reasoning cites this paper.

Are Tools Always Beneficial? Learning to Invoke Tools Adaptively for Dual-Mode Multimodal LLM Reasoning G-LLaVA: Solving Geometric Problem with Multi-Modal Large Language Model

Reference 16

Resolution
metadata mismatch
arxiv_id, observed 2026-05-20T06:18:05.269123Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-05-20T06:16:47.650748Z digest=sha256:9b645c89db5d341bf8cb3848480c95b71fd5dfa93a5f5137970af832646b6fe7

Observation 5715ec37-505e-47c5-b169-dff5fe7e45ab · inbound

From Seeing to Thinking: Decoupling Perception and Reasoning Improves Post-Training of Vision-Language Models cites this paper.

From Seeing to Thinking: Decoupling Perception and Reasoning Improves Post-Training of Vision-Language Models G-LLaVA: Solving Geometric Problem with Multi-Modal Large Language Model

Reference 5

Resolution
verified exact
arxiv_id, observed 2026-05-20T05:13:21.504876Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-20T05:13:03.237427Z digest=sha256:12ad82681813ac08c7dafaa1baf14112eeacd50da96eb9ae8e950e97028faed4

Observation 6af8c992-fe33-443f-9842-87ca4f8fece6 · inbound

Zamba2-VL Technical Report cites this paper.

Zamba2-VL Technical Report G-LLaVA: Solving Geometric Problem with Multi-Modal Large Language Model

Reference 135

Resolution
verified exact
arxiv_id, observed 2026-07-01T19:26:00.656851Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-06-28T22:34:20.970856Z digest=sha256:9315f58f0c98e17ac8258a5acf81bc0dec21596d30f7a8e94339dccf8bf182ac

Observation a8895596-8416-4bb2-ba98-eb44d5fd8305 · inbound

TRON: Targeted Rule-Verifiable Online Environments for Visual Reasoning RL cites this paper.

TRON: Targeted Rule-Verifiable Online Environments for Visual Reasoning RL G-LLaVA: Solving Geometric Problem with Multi-Modal Large Language Model

Reference 10

Resolution
verified exact
arxiv_id, observed 2026-07-01T22:56:19.953715Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-06-28T14:55:07.045625Z digest=sha256:3e7a0db41f888b470b7afaab94b124420e61efd0dc83a4b182de4da3c46cbbc4

Observation c9709043-4672-417b-9a3c-a28ee2293e47 · inbound

BiNSGPS: Geometry Problem Solving via Bidirectional Neuro-Symbolic Interaction cites this paper.

BiNSGPS: Geometry Problem Solving via Bidirectional Neuro-Symbolic Interaction G-LLaVA: Solving Geometric Problem with Multi-Modal Large Language Model

Reference 19

Resolution
metadata mismatch
arxiv_id, observed 2026-07-02T08:16:47.397677Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-06-28T06:18:38.325353Z digest=sha256:c38d76e270977da1493e7ad6c1ff5c68e93bb22cafa2420c452cf3b06f88c457

Observation a7dd9271-5588-40d9-950c-45f4726a6b0f · inbound

Anchored, Not Graded: Vision-Language Models Fail at Slant-from-Texture Perception cites this paper.

Anchored, Not Graded: Vision-Language Models Fail at Slant-from-Texture Perception G-LLaVA: Solving Geometric Problem with Multi-Modal Large Language Model

Reference 9

Resolution
metadata mismatch
arxiv_id, observed 2026-07-02T13:06:59.534118Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-06-28T01:33:01.158026Z digest=sha256:632b37155d35fcbc56fde8836e8f49e733a7da226ca348ca93f0436318490321

Observation 5cb66584-f6b5-4bb8-b2c3-4ecfecaf2018 · inbound

Anchored, Not Graded: Vision-Language Models Fail at Slant-from-Texture Perception cites this paper.

Anchored, Not Graded: Vision-Language Models Fail at Slant-from-Texture Perception G-LLaVA: Solving Geometric Problem with Multi-Modal Large Language Model

Reference 11

Resolution
metadata mismatch
arxiv_id, observed 2026-07-02T22:57:25.496195Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-07-02T22:56:50.330425Z digest=sha256:d20dc86b1c8103c99ab2401b57da34cbe2a1560ad0fc30687c09990d08ac5191

Observation 79ec3091-6e26-40d5-8202-80f0c3944fb2 · inbound

Closed-Form Spectral Regularization for Multi-Task Model Merging cites this paper.

Closed-Form Spectral Regularization for Multi-Task Model Merging G-LLaVA: Solving Geometric Problem with Multi-Modal Large Language Model

Reference 79

Resolution
verified exact
arxiv_id, observed 2026-07-02T16:27:09.392851Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-06-27T22:40:00.510742Z digest=sha256:88803153686163326528f05ba13cf85930d312d81f52475cb1b6808980dba73d

Observation 80ae989d-a7dd-4bf1-a1eb-6e12f64bdc2e · inbound

Artificial Intelligence for Mathematical Reasoning: An Integrated Survey of Language Models, Neuro-symbolic Systems, and Verified Discovery cites this paper.

Artificial Intelligence for Mathematical Reasoning: An Integrated Survey of Language Models, Neuro-symbolic Systems, and Verified Discovery G-LLaVA: Solving Geometric Problem with Multi-Modal Large Language Model

Reference 36

Resolution
verified exact
arxiv_id, observed 2026-07-02T22:47:26.050754Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-06-27T18:39:44.696961Z digest=sha256:690fa92219306de37a0531461dbd6da7abd67cccac7965617aa87fae782a9e1b

Observation 3359b26d-1e99-49e0-b752-bf02c2d24cd1 · inbound

Artificial Intelligence for Mathematical Reasoning: An Integrated Survey of Language Models, Neuro-symbolic Systems, and Verified Discovery cites this paper.

Artificial Intelligence for Mathematical Reasoning: An Integrated Survey of Language Models, Neuro-symbolic Systems, and Verified Discovery G-LLaVA: Solving Geometric Problem with Multi-Modal Large Language Model

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-02T12:05:04.665244Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T12:05:04.665244Z digest=sha256:5c515e75ae82bf6f4f9220ed0f90cb4ed203114d4e8f808e2997b5fa244cefe5

Observation 7726e9b1-b2bd-4dff-84e5-570f046ed286 · inbound

MathVis-Fine: Aligning Visual Supervision with Necessity via Progressive Dependency-Guided Training for Multimodal Mathematical Reasoning cites this paper.

MathVis-Fine: Aligning Visual Supervision with Necessity via Progressive Dependency-Guided Training for Multimodal Mathematical Reasoning G-LLaVA: Solving Geometric Problem with Multi-Modal Large Language Model

Reference 43

Resolution
metadata mismatch
arxiv_id, observed 2026-07-03T20:18:57.782252Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-06-27T01:23:40.564561Z digest=sha256:af30e3bffd713d15888d8de07c6ea990f9456c58db8f1dd89910f803a8b6e422

Observation 854589ce-debe-4e2e-9ac8-5f2409d10919 · inbound

Zone of Proximal Policy Optimization: Teacher in Prompts, Not Gradients cites this paper.

Zone of Proximal Policy Optimization: Teacher in Prompts, Not Gradients G-LLaVA: Solving Geometric Problem with Multi-Modal Large Language Model

Reference 98

Resolution
verified exact
arxiv_id, observed 2026-07-03T20:48:56.187511Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-06-27T01:08:52.981296Z digest=sha256:8b60ae2e2448e70388ac5341bd5d78cf967d742d86190b8121300f831f723d62

Observation 196d9c93-8c4d-4a1b-9a98-38781c31278a · inbound

Latent Visual States for Efficient Multimodal Reasoning cites this paper.

Latent Visual States for Efficient Multimodal Reasoning G-LLaVA: Solving Geometric Problem with Multi-Modal Large Language Model

Reference 10

Resolution
verified exact
arxiv_id, observed 2026-07-04T16:29:57.270529Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-06-26T00:38:11.619574Z digest=sha256:c6bd3765c47a55a843b4926b2fae21ecf35601620e12ea1d5d60fd8a8465ab3b

Observation 9c5d7403-a04f-43d1-b567-b3c9a1805600 · inbound

OpenCoF: Learning to Reason Through Video Generation cites this paper.

OpenCoF: Learning to Reason Through Video Generation G-LLaVA: Solving Geometric Problem with Multi-Modal Large Language Model

Reference 7

Resolution
verified exact
local_arxiv, observed 2026-07-10T01:46:40.940154Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-07-10T01:43:37.265723Z digest=sha256:1905fde5b28e19e31ac145a96804c274e27658e8e61742bc1fe256b5c380425c

Observation f1775b09-2046-45cd-b9f9-86a2151b095c · inbound

HalluScope: Fine-grained Hallucination Diagnosis for Multimodal Large Language Models cites this paper.

HalluScope: Fine-grained Hallucination Diagnosis for Multimodal Large Language Models G-LLaVA: Solving Geometric Problem with Multi-Modal Large Language Model

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-01T08:33:07.188782Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T08:33:07.188782Z digest=sha256:fc872831d67804e3729747166d9e7bdabe5bee4dc758bb59b62f06ef40c3e568

Observation 40db45d6-95b0-417b-bd60-f95f7464cac1 · inbound

Self-Boosting Vision-Language Models with Noisy Student On-Policy Self-Distillation cites this paper.

Self-Boosting Vision-Language Models with Noisy Student On-Policy Self-Distillation G-LLaVA: Solving Geometric Problem with Multi-Modal Large Language Model

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-01T03:38:10.677757Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T03:38:10.677757Z digest=sha256:4fc2f2e7913a068f12072133670152440aa3d36643b1055b2c09385313da4997

Observation f30a77f1-38e0-4574-90c6-3e2c69dcff9e · inbound

Self-Boosting Vision-Language Models with Noisy Student On-Policy Self-Distillation cites this paper.

Self-Boosting Vision-Language Models with Noisy Student On-Policy Self-Distillation G-LLaVA: Solving Geometric Problem with Multi-Modal Large Language Model

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-03T01:52:57.247904Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T01:52:57.247904Z digest=sha256:c9c34d2c3aa9dc43b455a9440107c13aac2171469a8ab1155082be689162ba8d