Pith. sign in

Paper Citation Record · LEDGER

GLM-4.5V and GLM-4.1V-Thinking: Towards Versatile Multimodal Reasoning with Scalable Reinforcement Learning

As of 1 August 2026, this Paper Citation Record lists 77 of 77 outbound references and 100 inbound Pith citation observations for arXiv:2507.01006.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2507.01006 v6

Coverage vector

measured 77 of 77 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-05-11T04:48:26.355351Z

measured 177 of 177 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-01T06:32:01.292127+00:00

measured 100 of 186 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-07-14T19:27:29.843866Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-07-10T12:15:01.137692Z

Reference resolution

77 of 77 outbound references displayed

  • verified exact39
  • verified fuzzy10
  • unresolved26
  • parse uncertain1
  • malformed identifier1
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 0f95b7c9-4687-4c26-a1b3-b07c51f0fde0 · outbound

This paper cites an unresolved cited work.

GLM-4.5V and GLM-4.1V-Thinking: Towards Versatile Multimodal Reasoning with Scalable Reinforcement Learning Unresolved cited work

Reference 1

Resolution
unresolved
raw_fallback, observed 2026-05-11T04:48:27.323386Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.

source=pdf_text observed=2026-05-11T04:48:26.355351Z digest=sha256:d73b2468d43d661284c65a4358336a4af404a265057df1a3e168279ad460c205

Observation 9ddcf976-45af-4baa-a121-2c5ed0e47f5b · outbound

This paper cites an unresolved cited work.

GLM-4.5V and GLM-4.1V-Thinking: Towards Versatile Multimodal Reasoning with Scalable Reinforcement Learning Unresolved cited work

Reference 2

Resolution
unresolved
raw_fallback, observed 2026-05-11T04:48:27.276647Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.

source=pdf_text observed=2026-05-11T04:48:26.355351Z digest=sha256:7c1884e11ab6a4e82193a631e1a20f01557485c11407181e56fcd4de3fc489e9

Observation 0703c035-12af-430d-add3-ed115c692077 · outbound

This paper cites Awadalla, L.

GLM-4.5V and GLM-4.1V-Thinking: Towards Versatile Multimodal Reasoning with Scalable Reinforcement Learning Awadalla, L

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-05-11T04:48:27.292864Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.

source=pdf_text observed=2026-05-11T04:48:26.355351Z digest=sha256:268e1c27f1a8305cb2b5b34e75abd38222df9f8bf06c2c3fe60ef5c2dbb762a9

Observation 6d2b3236-2cc7-4962-a37f-f3d3f52ed4e9 · outbound

This paper cites Qwen2.5-VL Technical Report.

GLM-4.5V and GLM-4.1V-Thinking: Towards Versatile Multimodal Reasoning with Scalable Reinforcement Learning Qwen2.5-VL Technical Report

Reference 4

Resolution
verified exact
local_arxiv, observed 2026-05-11T04:48:26.625358Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.

source=pdf_text observed=2026-05-11T04:48:26.355351Z digest=sha256:acdad660784b5e40e43a9a12fae2ee31718c2381f02bce999f8b310e8a895043

Observation 50bb9f86-3afe-4d1d-9b76-aea2ab1ed454 · outbound

This paper cites Nougat: Neural Optical Understanding for Academic Documents.

GLM-4.5V and GLM-4.1V-Thinking: Towards Versatile Multimodal Reasoning with Scalable Reinforcement Learning Nougat: Neural Optical Understanding for Academic Documents

Reference 5

Resolution
verified exact
arxiv_id, observed 2026-05-16T09:42:12.674341Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.

source=pdf_text observed=2026-05-11T04:48:26.355351Z digest=sha256:7158565e9e25c5e25adb6e13569343a09dda7f8369e5f498d0c1652b528141fc

Observation f180e9f8-0e01-4774-891a-c5d27eb118ff · outbound

This paper cites an unresolved cited work.

GLM-4.5V and GLM-4.1V-Thinking: Towards Versatile Multimodal Reasoning with Scalable Reinforcement Learning Unresolved cited work

Reference 6

Resolution
unresolved
raw_fallback, observed 2026-05-11T04:48:27.346351Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.

source=pdf_text observed=2026-05-11T04:48:26.355351Z digest=sha256:f25955c28f97481a3cb532bd605953b65e0a7265356ef46582639d9b86e12787

Observation 2cd634ac-3e27-48dd-8d39-c6843befde26 · outbound

This paper cites Are We on the Right Way for Evaluating Large Vision-Language Models?.

GLM-4.5V and GLM-4.1V-Thinking: Towards Versatile Multimodal Reasoning with Scalable Reinforcement Learning Are We on the Right Way for Evaluating Large Vision-Language Models?

Reference 7

Resolution
verified exact
arxiv_id, observed 2026-05-12T19:41:44.612219Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.

source=pdf_text observed=2026-05-11T04:48:26.355351Z digest=sha256:989e88e86e2993cbb5d33df45e3de41201a91c080eee5f114d70c936a37983a9

Observation 13e17e04-725d-47c8-af5e-d63814117f46 · outbound

This paper cites Data Filtering Networks.

GLM-4.5V and GLM-4.1V-Thinking: Towards Versatile Multimodal Reasoning with Scalable Reinforcement Learning Data Filtering Networks

Reference 8

Resolution
verified exact
arxiv_id, observed 2026-05-11T04:48:26.707521Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.

source=pdf_text observed=2026-05-11T04:48:26.355351Z digest=sha256:89041345011379dbb2d1113cd8ddadf617603242203654815f6eb905bba7f056

Observation 01ba4a1b-7329-4088-ae27-c3d4ce0475f2 · outbound

This paper cites an unresolved cited work.

GLM-4.5V and GLM-4.1V-Thinking: Towards Versatile Multimodal Reasoning with Scalable Reinforcement Learning Unresolved cited work

Reference 9

Resolution
unresolved
raw_fallback, observed 2026-05-11T04:48:27.375785Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.

source=pdf_text observed=2026-05-11T04:48:26.355351Z digest=sha256:94a16ed62066b12b8322dd0e55c527af6a9baf63a2166f28d13481e6167fbe8c

Observation 12c0ffd9-cdca-4e23-9fc7-c72bae045066 · outbound

This paper cites Video-MME: The First-Ever Comprehensive Evaluation Benchmark of Multi-modal LLMs in Video Analysis.

GLM-4.5V and GLM-4.1V-Thinking: Towards Versatile Multimodal Reasoning with Scalable Reinforcement Learning Video-MME: The First-Ever Comprehensive Evaluation Benchmark of Multi-modal LLMs in Video Analysis

Reference 10

Resolution
verified exact
local_arxiv, observed 2026-05-11T04:48:26.728015Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.

source=pdf_text observed=2026-05-11T04:48:26.355351Z digest=sha256:f24bc82474a817da806a8d04f001ca79cce31547542d40339029905a7689a5d4

Observation 7095c77d-5117-459b-ac10-881acf5f2746 · outbound

This paper cites BLINK: Multimodal Large Language Models Can See but Not Perceive.

GLM-4.5V and GLM-4.1V-Thinking: Towards Versatile Multimodal Reasoning with Scalable Reinforcement Learning BLINK: Multimodal Large Language Models Can See but Not Perceive

Reference 11

Resolution
verified exact
arxiv_id, observed 2026-05-15T20:18:15.850579Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.

source=pdf_text observed=2026-05-11T04:48:26.355351Z digest=sha256:bdaee196940d37b6acb09b491f5040c895aab1ac57dddfbf2b0de12736c912a1

Observation 11e6216c-b098-4349-99e9-4960ddd5230c · outbound

This paper cites an unresolved cited work.

GLM-4.5V and GLM-4.1V-Thinking: Towards Versatile Multimodal Reasoning with Scalable Reinforcement Learning Unresolved cited work

Reference 12

Resolution
unresolved
raw_fallback, observed 2026-05-11T04:48:27.406264Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.

source=pdf_text observed=2026-05-11T04:48:26.355351Z digest=sha256:b8d12ac2b80eaab2a26ee7aaa7b6799585542bf56343e69bae687e4e159830e6

Observation 730bd810-dbe5-4bf2-99fe-2bc6d8e27692 · outbound

This paper cites ChatGLM: A Family of Large Language Models from GLM-130B to GLM-4 All Tools.

GLM-4.5V and GLM-4.1V-Thinking: Towards Versatile Multimodal Reasoning with Scalable Reinforcement Learning ChatGLM: A Family of Large Language Models from GLM-130B to GLM-4 All Tools

Reference 13

Resolution
verified exact
arxiv_id, observed 2026-05-11T08:08:10.048088Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.

source=pdf_text observed=2026-05-11T04:48:26.355351Z digest=sha256:5a6cc278b45cfcb5b2605b8d60ba54240c4465b63107d2d362bb80689bc48129

Observation 806904d4-06b1-4e3b-b08b-0a7c71fad7bb · outbound

This paper cites an unresolved cited work.

GLM-4.5V and GLM-4.1V-Thinking: Towards Versatile Multimodal Reasoning with Scalable Reinforcement Learning Unresolved cited work

Reference 14

Resolution
unresolved
raw_fallback, observed 2026-05-11T04:48:27.418528Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.

source=pdf_text observed=2026-05-11T04:48:26.355351Z digest=sha256:4fd9554a206ecb3bf8ccb9a6ac5f1c8377c29c2e6e5820f733dc3b2f12365c6d

Observation 6d70cd75-b56c-411b-be2f-2fb4c992aef5 · outbound

This paper cites an unresolved cited work.

GLM-4.5V and GLM-4.1V-Thinking: Towards Versatile Multimodal Reasoning with Scalable Reinforcement Learning Unresolved cited work

Reference 15

Resolution
unresolved
raw_fallback, observed 2026-05-11T04:48:27.430669Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.

source=pdf_text observed=2026-05-11T04:48:26.355351Z digest=sha256:9642755e8b8305de32b21e4b7842bb2eecfabe433d66066294f0b26cfee07490

Observation 31960e2f-eab4-4db4-a130-b6b293d6ee4e · outbound

This paper cites Seed1.5-VL Technical Report.

GLM-4.5V and GLM-4.1V-Thinking: Towards Versatile Multimodal Reasoning with Scalable Reinforcement Learning Seed1.5-VL Technical Report

Reference 16

Resolution
verified exact
arxiv_id, observed 2026-05-11T05:26:06.820639Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.

source=pdf_text observed=2026-05-11T04:48:26.355351Z digest=sha256:7a39d82a0eb6389b184619b32fc8e77a2e48cfd35da446080a6960a6b0852c2c

Observation 55341280-c5b0-4020-bcc2-8295d0970e8e · outbound

This paper cites DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning.

GLM-4.5V and GLM-4.1V-Thinking: Towards Versatile Multimodal Reasoning with Scalable Reinforcement Learning DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning

Reference 17

Resolution
verified exact
local_arxiv, observed 2026-05-11T04:48:26.783532Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.

source=pdf_text observed=2026-05-11T04:48:26.355351Z digest=sha256:1ffdbf5bfbcdc093189e58c10a1d5b893e3a9e444ed7e99c477fc0ab7b6ca59c

Observation 64f82a00-3b07-4bab-a1c4-79c8507edf7e · outbound

This paper cites WebVoyager: Building an End-to-End Web Agent with Large Multimodal Models.

GLM-4.5V and GLM-4.1V-Thinking: Towards Versatile Multimodal Reasoning with Scalable Reinforcement Learning WebVoyager: Building an End-to-End Web Agent with Large Multimodal Models

Reference 18

Resolution
verified exact
arxiv_id, observed 2026-05-15T22:43:35.479641Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.

source=pdf_text observed=2026-05-11T04:48:26.355351Z digest=sha256:46496d36f739e9d73d477e29943a0f8a623d3af8dee2d6f6436deb24a4256f7e

Observation 1d0cbc65-7340-4f17-b5f0-116809f90350 · outbound

This paper cites Hong*, Y.

GLM-4.5V and GLM-4.1V-Thinking: Towards Versatile Multimodal Reasoning with Scalable Reinforcement Learning Hong*, Y

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-05-11T04:48:27.459608Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.

source=pdf_text observed=2026-05-11T04:48:26.355351Z digest=sha256:69ce4973e1fb899a37ca77a25c2ea4e9dd411b843a68bbdea73ba4e217c33048

Observation c18bd244-ec45-4b50-8fdd-e5e641b6b1d3 · outbound

This paper cites CogVLM2: Visual Language Models for Image and Video Understanding.

GLM-4.5V and GLM-4.1V-Thinking: Towards Versatile Multimodal Reasoning with Scalable Reinforcement Learning CogVLM2: Visual Language Models for Image and Video Understanding

Reference 20

Resolution
verified exact
arxiv_id, observed 2026-05-16T20:10:28.051835Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.

source=pdf_text observed=2026-05-11T04:48:26.355351Z digest=sha256:18fbcce7a2aef9542f34f742eb26c48ddc9172dc21aef90bcb40a43636bec627

Observation 8f84ff08-b70d-4ef6-832e-dd9cf65f9e5e · outbound

This paper cites an unresolved cited work.

GLM-4.5V and GLM-4.1V-Thinking: Towards Versatile Multimodal Reasoning with Scalable Reinforcement Learning Unresolved cited work

Reference 21

Resolution
unresolved
raw_fallback, observed 2026-05-11T04:48:27.476858Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.

source=pdf_text observed=2026-05-11T04:48:26.355351Z digest=sha256:3564fd3fbbf2dbd2d80459415cf458621b4a31c5da2b9717304a4e22476e1663

Observation efe7b67b-80d4-4485-a868-b065ed9921bb · outbound

This paper cites an unresolved cited work.

GLM-4.5V and GLM-4.1V-Thinking: Towards Versatile Multimodal Reasoning with Scalable Reinforcement Learning Unresolved cited work

Reference 22

Resolution
unresolved
raw_fallback, observed 2026-05-11T04:48:27.481408Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.

source=pdf_text observed=2026-05-11T04:48:26.355351Z digest=sha256:f36cf2855445cd3ebb89f4b56801e016517cceef5ff8b2355c96ac52f9f7943c

Observation 9de0b98c-a221-4670-9e94-57be5e67862a · outbound

This paper cites OpenAI o1 System Card.

GLM-4.5V and GLM-4.1V-Thinking: Towards Versatile Multimodal Reasoning with Scalable Reinforcement Learning OpenAI o1 System Card

Reference 23

Resolution
verified exact
local_arxiv, observed 2026-05-11T04:48:26.835382Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.

source=pdf_text observed=2026-05-11T04:48:26.355351Z digest=sha256:a2bf484471277d1f63293a60d0a1daed282a122ba374d5bec75c233cd2cdf19b

Observation b1d80984-08e1-4888-a774-275d110f7f57 · outbound

This paper cites Omnispatial: Towards comprehensive spatial reasoning benchmark for vision language models.

GLM-4.5V and GLM-4.1V-Thinking: Towards Versatile Multimodal Reasoning with Scalable Reinforcement Learning Omnispatial: Towards comprehensive spatial reasoning benchmark for vision language models

Reference 24

Resolution
verified exact
arxiv_id, observed 2026-05-11T04:48:26.845811Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.

source=pdf_text observed=2026-05-11T04:48:26.355351Z digest=sha256:a632297522727c91496769ee426081117c34b009ad6dfecaa8fe6b06d8e71d1c

Observation 32557da0-c972-449a-85bf-8ed1c30ca02b · outbound

This paper cites Kazemzadeh, V.

GLM-4.5V and GLM-4.1V-Thinking: Towards Versatile Multimodal Reasoning with Scalable Reinforcement Learning Kazemzadeh, V

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-05-11T04:48:27.503174Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.

source=pdf_text observed=2026-05-11T04:48:26.355351Z digest=sha256:2147cbe1fc70bb6528ba0b5bf48c82109f81c007af25b3dae7979bbfe0b9079c

Observation 993f66ea-087e-4945-bf97-538b2cd6099a · outbound

This paper cites Kembhavi, M.

GLM-4.5V and GLM-4.1V-Thinking: Towards Versatile Multimodal Reasoning with Scalable Reinforcement Learning Kembhavi, M

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-05-11T04:48:27.507506Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.

source=pdf_text observed=2026-05-11T04:48:26.355351Z digest=sha256:2a9556926940b6d428b43159031365a0065bb06972974142b52b857272a37673

Observation cbf6d70e-2a8d-42e5-9e97-42948754d08d · outbound

This paper cites an unresolved cited work.

GLM-4.5V and GLM-4.1V-Thinking: Towards Versatile Multimodal Reasoning with Scalable Reinforcement Learning Unresolved cited work

Reference 27

Resolution
unresolved
raw_fallback, observed 2026-05-11T04:48:27.518542Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.

source=pdf_text observed=2026-05-11T04:48:26.355351Z digest=sha256:8b702596ae6348e2f6b50a096d774a39b26894ab28c955f426a889791e6f8f2f

Observation 40cb6362-eb21-4cc3-9d62-be85682f8d94 · outbound

This paper cites an unresolved cited work.

GLM-4.5V and GLM-4.1V-Thinking: Towards Versatile Multimodal Reasoning with Scalable Reinforcement Learning Unresolved cited work

Reference 28

Resolution
unresolved
raw_fallback, observed 2026-05-11T04:48:27.522931Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.

source=pdf_text observed=2026-05-11T04:48:26.355351Z digest=sha256:a9e339a170c835e57576727177d114fd7ad34ea8019929789aa0c6def9ff6278

Observation 43d12450-2c33-4b38-841a-93432052401f · outbound

This paper cites OmniCorpus: A Unified Multimodal Corpus of 10 Billion-Level Images Interleaved with Text.

GLM-4.5V and GLM-4.1V-Thinking: Towards Versatile Multimodal Reasoning with Scalable Reinforcement Learning OmniCorpus: A Unified Multimodal Corpus of 10 Billion-Level Images Interleaved with Text

Reference 29

Resolution
verified exact
arxiv_id, observed 2026-05-11T04:48:26.861808Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.

source=pdf_text observed=2026-05-11T04:48:26.355351Z digest=sha256:286ff5593ec583c021b3da65739703861808c2c6a50eaa0449c9941701103f6c

Observation eb8a6873-00c2-44a1-ac9a-d943b31a1ae2 · outbound

This paper cites MMBench: Is Your Multi-modal Model an All-around Player?.

GLM-4.5V and GLM-4.1V-Thinking: Towards Versatile Multimodal Reasoning with Scalable Reinforcement Learning MMBench: Is Your Multi-modal Model an All-around Player?

Reference 30

Resolution
verified exact
arxiv_id, observed 2026-05-12T17:20:54.147488Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.

source=pdf_text observed=2026-05-11T04:48:26.355351Z digest=sha256:7e18b3614469e6a85b991ecd0be216d1fa2cf2275fbd7ec932e8da3cb6c13170

Observation 89b20bb9-f45e-416d-937f-80a89890e796 · outbound

This paper cites an unresolved cited work.

GLM-4.5V and GLM-4.1V-Thinking: Towards Versatile Multimodal Reasoning with Scalable Reinforcement Learning Unresolved cited work

Reference 31

Resolution
unresolved
raw_fallback, observed 2026-05-11T04:48:27.558472Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.

source=pdf_text observed=2026-05-11T04:48:26.355351Z digest=sha256:e670ff3ea57b3a78f519ca5559c23fafaab942a33b311bf7703c4428d105717b

Observation 1cb2b56c-d296-4890-b665-1d6162ec25f1 · outbound

This paper cites MathVista: Evaluating Mathematical Reasoning of Foundation Models in Visual Contexts.

GLM-4.5V and GLM-4.1V-Thinking: Towards Versatile Multimodal Reasoning with Scalable Reinforcement Learning MathVista: Evaluating Mathematical Reasoning of Foundation Models in Visual Contexts

Reference 32

Resolution
verified exact
local_arxiv, observed 2026-05-11T04:48:26.890437Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.

source=pdf_text observed=2026-05-11T04:48:26.355351Z digest=sha256:76ed6e9edd91bc3d17dd4a6f2abf5d9cea2cb65c55e1b994836b6c047f9dd07d

Observation a4b5b0b0-401f-4f6b-af39-858f318d383e · outbound

This paper cites an unresolved cited work.

GLM-4.5V and GLM-4.1V-Thinking: Towards Versatile Multimodal Reasoning with Scalable Reinforcement Learning Unresolved cited work

Reference 33

Resolution
unresolved
raw_fallback, observed 2026-05-11T04:48:27.576258Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.

source=pdf_text observed=2026-05-11T04:48:26.355351Z digest=sha256:ed63afa4cb1d504316133cb888703b9b82f7a2fceb54741d33496fd49e0237aa

Observation 38a407fc-7add-4f1b-ba85-90824b936f1c · outbound

This paper cites an unresolved cited work.

GLM-4.5V and GLM-4.1V-Thinking: Towards Versatile Multimodal Reasoning with Scalable Reinforcement Learning Unresolved cited work

Reference 34

Resolution
unresolved
raw_fallback, observed 2026-05-11T04:48:27.579852Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.

source=pdf_text observed=2026-05-11T04:48:26.355351Z digest=sha256:3443cae1772ab6053e96c7d27dd4a8c0e2043a8ab084207fb684f7f16aa219d8

Observation 431ac938-5fb6-42e3-9524-b12e1ed4f750 · outbound

This paper cites Masry, M.

GLM-4.5V and GLM-4.1V-Thinking: Towards Versatile Multimodal Reasoning with Scalable Reinforcement Learning Masry, M

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-05-11T04:48:27.588986Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.

source=pdf_text observed=2026-05-11T04:48:26.355351Z digest=sha256:1d65579546c4190dfd7393d28170ed1c5ab6945498962edf3199cafd6d60f259

Observation d6a31d02-7dc9-4414-afa3-02e917d7b4da · outbound

This paper cites an unresolved cited work.

GLM-4.5V and GLM-4.1V-Thinking: Towards Versatile Multimodal Reasoning with Scalable Reinforcement Learning Unresolved cited work

Reference 36

Resolution
parse uncertain
raw_fallback, observed 2026-05-11T04:48:27.593197Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.

source=pdf_text observed=2026-05-11T04:48:26.355351Z digest=sha256:d644d4521dbfdd0adcee8f9edcb483b28ed08707922b9ad2bbf98ec87b0b8ef5

Observation c46b98c3-879b-4007-8a46-c206711c38d9 · outbound

This paper cites We-Math: Does Your Large Multimodal Model Achieve Human-like Mathematical Reasoning?.

GLM-4.5V and GLM-4.1V-Thinking: Towards Versatile Multimodal Reasoning with Scalable Reinforcement Learning We-Math: Does Your Large Multimodal Model Achieve Human-like Mathematical Reasoning?

Reference 37

Resolution
verified exact
arxiv_id, observed 2026-05-16T21:55:41.398838Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.

source=pdf_text observed=2026-05-11T04:48:26.355351Z digest=sha256:c6c004ad664494044be40b23f998dbcd473df5add2e6f85401f5fc9a6ce17ec9

Observation 2423f61a-8242-4915-9692-a1e5fe5c5259 · outbound

This paper cites AndroidWorld: A Dynamic Benchmarking Environment for Autonomous Agents.

GLM-4.5V and GLM-4.1V-Thinking: Towards Versatile Multimodal Reasoning with Scalable Reinforcement Learning AndroidWorld: A Dynamic Benchmarking Environment for Autonomous Agents

Reference 38

Resolution
verified exact
arxiv_id, observed 2026-05-13T12:06:13.928391Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.

source=pdf_text observed=2026-05-11T04:48:26.355351Z digest=sha256:0f6c15f892d7146f87ab9fa2553bcc1e706bcbaf0fb56d2eb71b6bc2761ec049

Observation 72ce693f-81b5-4041-8af2-6027aa1b2613 · outbound

This paper cites ZeroBench: An Impossible Visual Benchmark for Contemporary Large Multimodal Models.

GLM-4.5V and GLM-4.1V-Thinking: Towards Versatile Multimodal Reasoning with Scalable Reinforcement Learning ZeroBench: An Impossible Visual Benchmark for Contemporary Large Multimodal Models

Reference 39

Resolution
verified exact
arxiv_id, observed 2026-07-08T01:19:04.778549Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.

source=pdf_text observed=2026-05-11T04:48:26.355351Z digest=sha256:20ed1a0b1528b61439e3e3f0d415ec3da13249a7b8d6c870d48c486224e408e1

Observation d92aa059-df7b-4b4a-8fe9-d8e4fd31c8ce · outbound

This paper cites Schuhmann, R.

GLM-4.5V and GLM-4.1V-Thinking: Towards Versatile Multimodal Reasoning with Scalable Reinforcement Learning Schuhmann, R

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-05-11T04:48:27.304629Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.

source=pdf_text observed=2026-05-11T04:48:26.355351Z digest=sha256:bc66f3f7df45e5fc52b21d8c901db3fdc8606d00f58bd6d54af7d944fdc86bc1

Observation 349ccc8e-5008-48bf-8d04-fdc34e30e3e7 · outbound

This paper cites Spurious Rewards: Rethinking Training Signals in RLVR.

GLM-4.5V and GLM-4.1V-Thinking: Towards Versatile Multimodal Reasoning with Scalable Reinforcement Learning Spurious Rewards: Rethinking Training Signals in RLVR

Reference 41

Resolution
verified exact
arxiv_id, observed 2026-05-16T13:37:51.161107Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.

source=pdf_text observed=2026-05-11T04:48:26.355351Z digest=sha256:4794b41cdf38ff034ec3566ae215e53a759f890767bf2512f83a2e352585338a

Observation 74767c3a-3163-48f9-aa61-6eb3787ed9fa · outbound

This paper cites DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models.

GLM-4.5V and GLM-4.1V-Thinking: Towards Versatile Multimodal Reasoning with Scalable Reinforcement Learning DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models

Reference 42

Resolution
verified exact
local_arxiv, observed 2026-05-11T04:48:26.951495Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.

source=pdf_text observed=2026-05-11T04:48:26.355351Z digest=sha256:fd4e37ab64dbc959bbe96d6b8d1c351ceb1e93abc37358c91d30016adc09ac63

Observation ee73d3ff-7448-4b1f-9069-8da0443b4962 · outbound

This paper cites an unresolved cited work.

GLM-4.5V and GLM-4.1V-Thinking: Towards Versatile Multimodal Reasoning with Scalable Reinforcement Learning Unresolved cited work

Reference 43

Resolution
unresolved
raw_fallback, observed 2026-05-11T04:48:27.369166Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.

source=pdf_text observed=2026-05-11T04:48:26.355351Z digest=sha256:0ac18228f6a575646ddb24b23ed81a2f476aaeb1097e90cbba1f8ef341e68ca3

Observation 78859dd4-c892-47c9-84f7-1899d66dd444 · outbound

This paper cites RoFormer: Enhanced Transformer with Rotary Position Embedding.

GLM-4.5V and GLM-4.1V-Thinking: Towards Versatile Multimodal Reasoning with Scalable Reinforcement Learning RoFormer: Enhanced Transformer with Rotary Position Embedding

Reference 44

Resolution
verified exact
local_arxiv, observed 2026-05-11T04:48:26.970102Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.

source=pdf_text observed=2026-05-11T04:48:26.355351Z digest=sha256:c60fc7403da22d0e578aa2d64cc02e92f0b2d3b2b538514b286905f1201729f3

Observation 51038012-181e-4453-988b-970b0d93333a · outbound

This paper cites Chartmuseum: Testing visual reasoning capabilities of large vision-language models.

GLM-4.5V and GLM-4.1V-Thinking: Towards Versatile Multimodal Reasoning with Scalable Reinforcement Learning Chartmuseum: Testing visual reasoning capabilities of large vision-language models

Reference 45

Resolution
verified exact
arxiv_id, observed 2026-05-11T04:48:26.985869Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.

source=pdf_text observed=2026-05-11T04:48:26.355351Z digest=sha256:a8039458812b92c744b205d80064f35f7b6a0a0f3a92535d71bed3c513bb4400

Observation 6d2f40ed-dc13-4a01-8020-8e3dccc62377 · outbound

This paper cites an unresolved cited work.

GLM-4.5V and GLM-4.1V-Thinking: Towards Versatile Multimodal Reasoning with Scalable Reinforcement Learning Unresolved cited work

Reference 46

Resolution
unresolved
raw_fallback, observed 2026-05-11T04:48:27.411372Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.

source=pdf_text observed=2026-05-11T04:48:26.355351Z digest=sha256:3e32b72c99e8d701f96b00cbf51e193603b3fe31814f05a356e576b942b16044

Observation 09574fcb-d1d6-40e7-ac81-d4a098481613 · outbound

This paper cites an unresolved cited work.

GLM-4.5V and GLM-4.1V-Thinking: Towards Versatile Multimodal Reasoning with Scalable Reinforcement Learning Unresolved cited work

Reference 47

Resolution
unresolved
raw_fallback, observed 2026-05-11T04:48:27.438756Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.

source=pdf_text observed=2026-05-11T04:48:26.355351Z digest=sha256:4f675d95ec9672e7d26807215d9623fa9ccf75cfd6a034e52f76844fdf0bb333

Observation d8f5720a-9f83-41fc-bf08-1a70ebf61992 · outbound

This paper cites Gemma 3 Technical Report.

GLM-4.5V and GLM-4.1V-Thinking: Towards Versatile Multimodal Reasoning with Scalable Reinforcement Learning Gemma 3 Technical Report

Reference 48

Resolution
verified exact
local_arxiv, observed 2026-05-11T04:48:27.005806Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.

source=pdf_text observed=2026-05-11T04:48:26.355351Z digest=sha256:e96bab64b5f1babde646f38d462409216b5323f4e487456a07f5a0871c10c008

Observation bf6c33c7-b45b-4f61-80c0-78a734d9e414 · outbound

This paper cites Gemini Robotics: Bringing AI into the Physical World.

GLM-4.5V and GLM-4.1V-Thinking: Towards Versatile Multimodal Reasoning with Scalable Reinforcement Learning Gemini Robotics: Bringing AI into the Physical World

Reference 49

Resolution
verified exact
arxiv_id, observed 2026-05-11T15:30:26.569831Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.

source=pdf_text observed=2026-05-11T04:48:26.355351Z digest=sha256:39285eca043ab3ec41ee77e6ade3ff857e11e39d5ec14ee7804aa3a50df33a99

Observation 3a040154-0e21-4640-9cb5-e8dd365ed4c4 · outbound

This paper cites an unresolved cited work.

GLM-4.5V and GLM-4.1V-Thinking: Towards Versatile Multimodal Reasoning with Scalable Reinforcement Learning Unresolved cited work

Reference 50

Resolution
unresolved
raw_fallback, observed 2026-05-11T04:48:27.470462Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.

source=pdf_text observed=2026-05-11T04:48:26.355351Z digest=sha256:666350494303726a126cf5e75349090fa79217e3da8b24be7a68593d2e6df00c

Observation 950f6f15-60c6-459c-bbf0-058c428440c0 · outbound

This paper cites an unresolved cited work.

GLM-4.5V and GLM-4.1V-Thinking: Towards Versatile Multimodal Reasoning with Scalable Reinforcement Learning Unresolved cited work

Reference 51

Resolution
unresolved
raw_fallback, observed 2026-05-11T04:48:27.487978Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.

source=pdf_text observed=2026-05-11T04:48:26.355351Z digest=sha256:92906a44161e2e4030385c6d3daec8bd7b847ac5924cf00863325e37c8c096b6

Observation 16ebb76e-e919-452b-bb33-3967c54d334d · outbound

This paper cites Step-3 is large yet affordable: Model-system co-design for cost-effective decoding.

GLM-4.5V and GLM-4.1V-Thinking: Towards Versatile Multimodal Reasoning with Scalable Reinforcement Learning Step-3 is large yet affordable: Model-system co-design for cost-effective decoding

Reference 52

Resolution
verified exact
arxiv_id, observed 2026-05-11T04:48:27.026664Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.

source=pdf_text observed=2026-05-11T04:48:26.355351Z digest=sha256:6c8304862298308265019419b58f6392a496953cfb9cedd19e47d2125dae0c59

Observation 296e7026-5aaf-438a-899a-bce379d37712 · outbound

This paper cites MuirBench: A Comprehensive Benchmark for Robust Multi-image Understanding.

GLM-4.5V and GLM-4.1V-Thinking: Towards Versatile Multimodal Reasoning with Scalable Reinforcement Learning MuirBench: A Comprehensive Benchmark for Robust Multi-image Understanding

Reference 53

Resolution
verified exact
arxiv_id, observed 2026-05-17T01:09:30.603679Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.

source=pdf_text observed=2026-05-11T04:48:26.355351Z digest=sha256:54378a4a6450dd47ccb83431619ce8d8f5233db8a3485647db6de14430655d9d

Observation 0d602a1f-e0d8-4649-9d56-a6cf1bae3443 · outbound

This paper cites Traceable evidence enhanced visual grounded reasoning: Evaluation and methodology.

GLM-4.5V and GLM-4.1V-Thinking: Towards Versatile Multimodal Reasoning with Scalable Reinforcement Learning Traceable evidence enhanced visual grounded reasoning: Evaluation and methodology

Reference 54

Resolution
verified exact
arxiv_id, observed 2026-05-11T04:48:27.050964Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.

source=pdf_text observed=2026-05-11T04:48:26.355351Z digest=sha256:3faff15fd92631d8583a74e587c0ff80bb59358c337058e5fb088f14da355860

Observation 85460c8a-acfa-42a0-8744-dac368308d64 · outbound

This paper cites Measuring Multimodal Mathematical Reasoning with MATH-Vision Dataset.

GLM-4.5V and GLM-4.1V-Thinking: Towards Versatile Multimodal Reasoning with Scalable Reinforcement Learning Measuring Multimodal Mathematical Reasoning with MATH-Vision Dataset

Reference 55

Resolution
verified exact
arxiv_id, observed 2026-05-17T20:45:37.909362Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.

source=pdf_text observed=2026-05-11T04:48:26.355351Z digest=sha256:a7772b381af8ea1f3fd6599cdc6a19a97c93268539043ff2d7a73dbd6b8b2d7e

Observation db32267a-c286-446c-9c0b-b04c51edd6cb · outbound

This paper cites an unresolved cited work.

GLM-4.5V and GLM-4.1V-Thinking: Towards Versatile Multimodal Reasoning with Scalable Reinforcement Learning Unresolved cited work

Reference 56

Resolution
unresolved
raw_fallback, observed 2026-05-11T04:48:27.601349Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.

source=pdf_text observed=2026-05-11T04:48:26.355351Z digest=sha256:b2cc58da928f6bb9c0fd67b7a691b28c51b92c4078a8f77be88cc6000e1bddd5

Observation 9f40eb1e-658c-46e0-b9de-4a42237d515e · outbound

This paper cites Qwen2-VL: Enhancing Vision-Language Model's Perception of the World at Any Resolution.

GLM-4.5V and GLM-4.1V-Thinking: Towards Versatile Multimodal Reasoning with Scalable Reinforcement Learning Qwen2-VL: Enhancing Vision-Language Model's Perception of the World at Any Resolution

Reference 57

Resolution
verified exact
local_arxiv, observed 2026-05-11T04:48:27.083792Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.

source=pdf_text observed=2026-05-11T04:48:26.355351Z digest=sha256:650dfd95169eecb2fd912ac0b32d6a295eeb3ac6dd3acd9fe5a3b908feaf5d7b

Observation 1a53ad19-5d81-47b2-ac86-96d4e2691dd5 · outbound

This paper cites LVBench: An Extreme Long Video Understanding Benchmark.

GLM-4.5V and GLM-4.1V-Thinking: Towards Versatile Multimodal Reasoning with Scalable Reinforcement Learning LVBench: An Extreme Long Video Understanding Benchmark

Reference 58

Resolution
verified exact
arxiv_id, observed 2026-05-19T11:55:30.239348Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.

source=pdf_text observed=2026-05-11T04:48:26.355351Z digest=sha256:f1c3fa61208ca9e5499f30cf1897cab2f4e2232b414b6ed38d20dbcae20a1f5c

Observation 5a1fbefe-20b0-4d5f-8a13-5f05abbdd271 · outbound

This paper cites CogVLM: Visual Expert for Pretrained Language Models.

GLM-4.5V and GLM-4.1V-Thinking: Towards Versatile Multimodal Reasoning with Scalable Reinforcement Learning CogVLM: Visual Expert for Pretrained Language Models

Reference 59

Resolution
verified exact
arxiv_id, observed 2026-05-15T15:46:06.589810Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.

source=pdf_text observed=2026-05-11T04:48:26.355351Z digest=sha256:18a0583fba91e0de5747759be19ac0d877a831f02cbd575a5045c01c3c24c9fb

Observation 4839ca8d-2669-4a28-9ef6-f690340e6c32 · outbound

This paper cites an unresolved cited work.

GLM-4.5V and GLM-4.1V-Thinking: Towards Versatile Multimodal Reasoning with Scalable Reinforcement Learning Unresolved cited work

Reference 60

Resolution
unresolved
raw_fallback, observed 2026-05-11T04:48:27.364215Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.

source=pdf_text observed=2026-05-11T04:48:26.355351Z digest=sha256:0d2db53e2bdfd634cf35cfa2ca90d9a4577677ce268694528f889b3ca871e34c

Observation 47da3101-6565-451c-a7f4-477a369b6024 · outbound

This paper cites an unresolved cited work.

GLM-4.5V and GLM-4.1V-Thinking: Towards Versatile Multimodal Reasoning with Scalable Reinforcement Learning Unresolved cited work

Reference 61

Resolution
unresolved
raw_fallback, observed 2026-05-11T04:48:27.384381Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.

source=pdf_text observed=2026-05-11T04:48:26.355351Z digest=sha256:cb3f127e16fc5c94789aedbfb627f82944a5359f5689fc93b19eaaa275800af6

Observation fb1afeb4-99da-4d24-9b96-30288b2de71b · outbound

This paper cites an unresolved cited work.

GLM-4.5V and GLM-4.1V-Thinking: Towards Versatile Multimodal Reasoning with Scalable Reinforcement Learning Unresolved cited work

Reference 62

Resolution
unresolved
raw_fallback, observed 2026-05-11T04:48:27.395274Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.

source=pdf_text observed=2026-05-11T04:48:26.355351Z digest=sha256:be64888040d9a8f559838f47ecb4410be1f852c1b798f8d848c17054247a85b4

Observation c2ca68de-de23-40c9-bf90-4f0d83d30cf3 · outbound

This paper cites Demystifying CLIP Data.

GLM-4.5V and GLM-4.1V-Thinking: Towards Versatile Multimodal Reasoning with Scalable Reinforcement Learning Demystifying CLIP Data

Reference 63

Resolution
verified exact
arxiv_id, observed 2026-05-16T09:20:20.637207Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.

source=pdf_text observed=2026-05-11T04:48:26.355351Z digest=sha256:c753d3438d51ddd248de5a37e03480079276d19a8e7fd3612d2562c0407ecadc

Observation 06fce0de-9d53-4c7f-9090-1838a236a2ae · outbound

This paper cites Scalable Chain of Thoughts via Elastic Reasoning.

GLM-4.5V and GLM-4.1V-Thinking: Towards Versatile Multimodal Reasoning with Scalable Reinforcement Learning Scalable Chain of Thoughts via Elastic Reasoning

Reference 64

Resolution
verified exact
arxiv_id, observed 2026-05-11T04:48:27.153354Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.

source=pdf_text observed=2026-05-11T04:48:26.355351Z digest=sha256:9b342c87320569a99448ac4770fc76a515e2c0d9d6b3c9735a9e9493607f9ece

Observation 7e952fe8-0505-4b58-9ece-17e1b33c6bed · outbound

This paper cites Seeing from Another Perspective: Evaluating Multi-View Understanding in MLLMs.

GLM-4.5V and GLM-4.1V-Thinking: Towards Versatile Multimodal Reasoning with Scalable Reinforcement Learning Seeing from Another Perspective: Evaluating Multi-View Understanding in MLLMs

Reference 65

Resolution
verified exact
arxiv_id, observed 2026-05-11T04:48:27.186831Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.

source=pdf_text observed=2026-05-11T04:48:26.355351Z digest=sha256:1cbfd6b601cef6c55228323e86d375d4e17719a4398fe906ee6e4324ea519e15

Observation a3d504d5-3782-45a7-81fd-467558ba43c8 · outbound

This paper cites DAPO: An Open-Source LLM Reinforcement Learning System at Scale.

GLM-4.5V and GLM-4.1V-Thinking: Towards Versatile Multimodal Reasoning with Scalable Reinforcement Learning DAPO: An Open-Source LLM Reinforcement Learning System at Scale

Reference 66

Resolution
verified exact
local_arxiv, observed 2026-05-11T04:48:27.221885Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.

source=pdf_text observed=2026-05-11T04:48:26.355351Z digest=sha256:2cfbdf26b3085c30540e700290ebc1fd62c1af062e11ea4d4f92a5515a9f75e9

Observation 5911b8df-ebd2-4ba7-82d1-5d46a91ce7e9 · outbound

This paper cites an unresolved cited work.

GLM-4.5V and GLM-4.1V-Thinking: Towards Versatile Multimodal Reasoning with Scalable Reinforcement Learning Unresolved cited work

Reference 67

Resolution
unresolved
raw_fallback, observed 2026-05-11T04:48:27.549350Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.

source=pdf_text observed=2026-05-11T04:48:26.355351Z digest=sha256:4aaef1c782cc2d5d02228823076b318adaa122d17b75fb0280df8657e0394d81

Observation 3964833b-e941-4864-abbe-307ba3d5b5ae · outbound

This paper cites MMMU-Pro: A More Robust Multi-discipline Multimodal Understanding Benchmark.

GLM-4.5V and GLM-4.1V-Thinking: Towards Versatile Multimodal Reasoning with Scalable Reinforcement Learning MMMU-Pro: A More Robust Multi-discipline Multimodal Understanding Benchmark

Reference 68

Resolution
verified exact
arxiv_id, observed 2026-05-14T00:51:48.633687Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.

source=pdf_text observed=2026-05-11T04:48:26.355351Z digest=sha256:ef8d1c37ae9b1e40f427b9a3301611c7a10b7a24d0a6c9d9e19c07e40de165dc

Observation 665bbba4-4de0-4785-b757-3c8c2f9fc193 · outbound

This paper cites Zhang, P.

GLM-4.5V and GLM-4.1V-Thinking: Towards Versatile Multimodal Reasoning with Scalable Reinforcement Learning Zhang, P

Reference 69

Resolution
verified fuzzy
raw_fallback, observed 2026-05-11T04:48:27.606139Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.

source=pdf_text observed=2026-05-11T04:48:26.355351Z digest=sha256:515bdf32a38a412dc1d51eaa507b550f3a49a00e2ef86b573f0e279755b9fba7

Observation 28db0541-c2dc-49fd-b387-434d2552bbbd · outbound

This paper cites Zhang, D.

GLM-4.5V and GLM-4.1V-Thinking: Towards Versatile Multimodal Reasoning with Scalable Reinforcement Learning Zhang, D

Reference 70

Resolution
verified fuzzy
raw_fallback, observed 2026-05-11T04:48:27.611199Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.

source=pdf_text observed=2026-05-11T04:48:26.355351Z digest=sha256:d58e241b0864bb71e5c6326e3854c91afdbfd617ba7e64239d066d02ed4a03bc

Observation 98d3a9b8-cef1-42f2-b348-c6ea1ea19948 · outbound

This paper cites an unresolved cited work.

GLM-4.5V and GLM-4.1V-Thinking: Towards Versatile Multimodal Reasoning with Scalable Reinforcement Learning Unresolved cited work

Reference 71

Resolution
unresolved
raw_fallback, observed 2026-05-11T04:48:27.445538Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.

source=pdf_text observed=2026-05-11T04:48:26.355351Z digest=sha256:f3065eb4638c86de34424b1461168e63412badb884458d4f418125621a115a97

Observation 22a1f710-e07d-4141-9ec1-bdb9ba20c755 · outbound

This paper cites InternVL3: Exploring Advanced Training and Test-Time Recipes for Open-Source Multimodal Models.

GLM-4.5V and GLM-4.1V-Thinking: Towards Versatile Multimodal Reasoning with Scalable Reinforcement Learning InternVL3: Exploring Advanced Training and Test-Time Recipes for Open-Source Multimodal Models

Reference 72

Resolution
verified exact
local_arxiv, observed 2026-05-11T04:48:27.253853Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.

source=pdf_text observed=2026-05-11T04:48:26.355351Z digest=sha256:077a514adfff925a309f41dabe7d2fbedd7b082cc42208549d137555cf90d64a

Observation e316bcbd-79bc-441a-9d9f-fd54c260487d · outbound

This paper cites an unresolved cited work.

GLM-4.5V and GLM-4.1V-Thinking: Towards Versatile Multimodal Reasoning with Scalable Reinforcement Learning Unresolved cited work

Reference 73

Resolution
unresolved
raw_fallback, observed 2026-05-11T04:48:27.496253Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.

source=pdf_text observed=2026-05-11T04:48:26.355351Z digest=sha256:2e0f2d56844a343fc1e92590844dfd99f34dc77c5be4389ad106759425c7f6bd

Observation b7f21d7b-1304-4d8e-aa23-099a7a46970d · outbound

This paper cites DynaMath: A Dynamic Visual Benchmark for Evaluating Mathematical Reasoning Robustness of Vision Language Models.

GLM-4.5V and GLM-4.1V-Thinking: Towards Versatile Multimodal Reasoning with Scalable Reinforcement Learning DynaMath: A Dynamic Visual Benchmark for Evaluating Mathematical Reasoning Robustness of Vision Language Models

Reference 74

Resolution
verified exact
arxiv_id, observed 2026-05-11T04:48:27.265524Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.

source=pdf_text observed=2026-05-11T04:48:26.355351Z digest=sha256:bbcd8f1284e790d9f4b587b4429a4b26b9999d25b05f4815e369089a986374d6

Observation 03b47436-6993-4564-99ed-0d555f18b7e6 · outbound

This paper cites 毛细管”。 当左右两个装有不同颜色液体的杯子与中间的空杯之间用纸巾连接时,纸巾会利用自身吸水性和纤维间的毛细作 用,将左侧红色液体和右侧蓝色液体通过纤维间隙输送至中间的空杯中。随着这种输送过程的进行,中间的空杯逐 渐被液体填满,从而出现了“中间水杯有水.

GLM-4.5V and GLM-4.1V-Thinking: Towards Versatile Multimodal Reasoning with Scalable Reinforcement Learning 毛细管”。 当左右两个装有不同颜色液体的杯子与中间的空杯之间用纸巾连接时,纸巾会利用自身吸水性和纤维间的毛细作 用,将左侧红色液体和右侧蓝色液体通过纤维间隙输送至中间的空杯中。随着这种输送过程的进行,中间的空杯逐 渐被液体填满,从而出现了“中间水杯有水

Reference 75

Resolution
verified fuzzy
raw_fallback, observed 2026-05-11T04:48:27.571807Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.

source=pdf_text observed=2026-05-11T04:48:26.355351Z digest=sha256:8c8695797e24afc0da358299b1330efaa6dec44e47eacb1e1bf8d2916895ec8c

Observation 29992f48-4b91-4465-a9f7-0ce5965452ca · outbound

This paper cites Meeting" event - October 9th has a.

GLM-4.5V and GLM-4.1V-Thinking: Towards Versatile Multimodal Reasoning with Scalable Reinforcement Learning Meeting" event - October 9th has a

Reference 76

Resolution
verified fuzzy
raw_fallback, observed 2026-05-11T04:48:27.452772Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.

source=pdf_text observed=2026-05-11T04:48:26.355351Z digest=sha256:79cd13dc02e403e5bf3de714b6efffa3cedb8dd0454d6d78e063903b07865f61

Observation 503ed9bb-710a-49d8-bfed-7baad49ef859 · outbound

This paper cites an unresolved cited work.

GLM-4.5V and GLM-4.1V-Thinking: Towards Versatile Multimodal Reasoning with Scalable Reinforcement Learning Unresolved cited work

Reference 77

Resolution
malformed identifier
raw_fallback, observed 2026-05-11T04:48:27.530294Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.

source=pdf_text observed=2026-05-11T04:48:26.355351Z digest=sha256:ba36636cc802e040e539fae5eb073a2e6551ba5eef892e6f313eeaee7fc839f9

Pith citing papers

Observation 2588ccd0-f7e3-4828-9f7f-5775d893a168 · inbound

LVBench: An Extreme Long Video Understanding Benchmark cites this paper.

LVBench: An Extreme Long Video Understanding Benchmark GLM-4.5V and GLM-4.1V-Thinking: Towards Versatile Multimodal Reasoning with Scalable Reinforcement Learning

Reference 13

Resolution
verified exact
local_arxiv, observed 2026-05-19T11:55:30.179227Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.

source=pdf_text observed=2026-05-19T11:55:30.048525Z digest=sha256:5a704d2ef057cdc8cb1ae89d08da22d0ef4824c28db7d5484f4d23395763607d

Observation 70b50022-3abd-4c81-bea8-5d283a80d8d0 · inbound

InternVL3.5: Advancing Open-Source Multimodal Models in Versatility, Reasoning, and Efficiency cites this paper.

InternVL3.5: Advancing Open-Source Multimodal Models in Versatility, Reasoning, and Efficiency GLM-4.5V and GLM-4.1V-Thinking: Towards Versatile Multimodal Reasoning with Scalable Reinforcement Learning

Reference 46

Resolution
verified exact
arxiv_id, observed 2026-05-11T04:48:27.616218Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.

source=pdf_text observed=2026-05-10T11:58:58.660564Z digest=sha256:1431e9910e3c77c9ad1921c42d6c1a112f2c164e0027eeeb58fc6c353aff7dcb

Observation 9118215b-73d7-4e4b-bfa6-04ccd9dc88fe · inbound

VeriOS: Query-Driven Proactive Human-Agent-GUI Interaction for Trustworthy OS Agents cites this paper.

VeriOS: Query-Driven Proactive Human-Agent-GUI Interaction for Trustworthy OS Agents GLM-4.5V and GLM-4.1V-Thinking: Towards Versatile Multimodal Reasoning with Scalable Reinforcement Learning

Reference 46

Resolution
verified exact
local_arxiv, observed 2026-05-18T18:06:42.564542Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.

source=pdf_text observed=2026-05-18T18:06:12.349285Z digest=sha256:76e08874fdd8368f24b55bb78a301d5f71e5675e4ec853bee0bcada8c07a1bfe

Observation 4a8bdd1e-5291-4f21-a081-9e8b64bd5e74 · inbound

RepIt: Steering Language Models with Concept-Specific Refusal Vectors cites this paper.

RepIt: Steering Language Models with Concept-Specific Refusal Vectors GLM-4.5V and GLM-4.1V-Thinking: Towards Versatile Multimodal Reasoning with Scalable Reinforcement Learning

Reference 5

Resolution
metadata mismatch
local_arxiv, observed 2026-05-18T16:06:34.627607Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.

source=pdf_text observed=2026-07-11T11:50:26.030339Z digest=sha256:c8e7a3513b83dfe9724be31e8f8d808413bdf82d8e5b70520557e8ef108b7bac

Observation f159faf4-89ee-4404-8c04-950a21f329ee · inbound

PiERN: Token-Level Routing for Integrating High-Precision Computation and Reasoning cites this paper.

PiERN: Token-Level Routing for Integrating High-Precision Computation and Reasoning GLM-4.5V and GLM-4.1V-Thinking: Towards Versatile Multimodal Reasoning with Scalable Reinforcement Learning

Reference 13

Resolution
metadata mismatch
local_arxiv, observed 2026-05-18T16:06:34.878621Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.

source=pdf_text observed=2026-05-18T16:05:57.505355Z digest=sha256:51bf4faee7f14110a54b61d7e7a581dadaaa14957aa46e6cd12811e333b5d8c8

Observation 0800ca9b-3731-44a1-b6c0-1a44082198da · inbound

SecureWebArena: A Holistic Security Evaluation Benchmark for LVLM-based Web Agents cites this paper.

SecureWebArena: A Holistic Security Evaluation Benchmark for LVLM-based Web Agents GLM-4.5V and GLM-4.1V-Thinking: Towards Versatile Multimodal Reasoning with Scalable Reinforcement Learning

Reference 38

Resolution
verified exact
local_arxiv, observed 2026-05-18T08:16:06.474138Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.

source=pdf_text observed=2026-05-18T08:14:51.102085Z digest=sha256:4e4dd9aa5125e6e0a0de192901e9d8f3a6281697f2e88324fd96c897a4ba36eb

Observation bd699a5d-3ef4-4bbd-9808-05ad9436106a · inbound

The Art of Scaling Reinforcement Learning Compute for LLMs cites this paper.

The Art of Scaling Reinforcement Learning Compute for LLMs GLM-4.5V and GLM-4.1V-Thinking: Towards Versatile Multimodal Reasoning with Scalable Reinforcement Learning

Reference 4

Resolution
metadata mismatch
local_arxiv, observed 2026-05-16T16:29:13.983934Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.

source=pdf_text observed=2026-05-16T16:29:13.954029Z digest=sha256:847828a83ed809759218309cae99ff686f763a6528ea71d9c3faf9b346d6a963

Observation d56bec6f-774d-4c95-9ee0-4f1f46841545 · inbound

LongVT: Incentivizing "Thinking with Long Videos" via Native Tool Calling cites this paper.

LongVT: Incentivizing "Thinking with Long Videos" via Native Tool Calling GLM-4.5V and GLM-4.1V-Thinking: Towards Versatile Multimodal Reasoning with Scalable Reinforcement Learning

Reference 12

Resolution
verified exact
local_arxiv, observed 2026-05-22T12:31:32.167022Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.

source=pdf_text observed=2026-05-22T12:26:35.347190Z digest=sha256:d658aec8caf13624f2fb5e54d5efbdba78a2caebf2fb65a94724d15880456411

Observation 0f9fd3c1-90cb-4d70-851b-db82c128c335 · inbound

Agentic Learner with Grow-and-Refine Multimodal Semantic Memory cites this paper.

Agentic Learner with Grow-and-Refine Multimodal Semantic Memory GLM-4.5V and GLM-4.1V-Thinking: Towards Versatile Multimodal Reasoning with Scalable Reinforcement Learning

Reference 29

Resolution
verified exact
local_arxiv, observed 2026-05-17T04:29:01.590494Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.

source=pdf_text observed=2026-05-17T04:27:40.232015Z digest=sha256:bbc5f22ca627889c0e7681a7c1e06c2c3669b15c2267eabc3103fecabdb0a55a

Observation c5e84e16-7e9a-49db-a209-f35e3638e387 · inbound

AgentProg: Empowering Long-Horizon GUI Agents with Program-Guided Context Management cites this paper.

AgentProg: Empowering Long-Horizon GUI Agents with Program-Guided Context Management GLM-4.5V and GLM-4.1V-Thinking: Towards Versatile Multimodal Reasoning with Scalable Reinforcement Learning

Reference 12

Resolution
verified exact
local_arxiv, observed 2026-05-16T23:43:42.344573Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.

source=pdf_text observed=2026-05-16T23:42:25.086602Z digest=sha256:32440a9745de93c575c73f08571f22fcc29b92101dd0fc6f813afba55507b78f

Observation e4860a6b-921b-4252-b950-e8b27b77849b · inbound

Setting the Stage: Text-Driven Scene-Consistent Image Generation cites this paper.

Setting the Stage: Text-Driven Scene-Consistent Image Generation GLM-4.5V and GLM-4.1V-Thinking: Towards Versatile Multimodal Reasoning with Scalable Reinforcement Learning

Reference 8

Resolution
verified exact
local_arxiv, observed 2026-05-21T17:44:17.309434Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.

source=pdf_text observed=2026-05-21T17:40:25.779794Z digest=sha256:b0594be45caed179380f1b9dcc92976893b76fdb282d2fa3ce53ac895cd7caf6

Observation a90fe606-918e-458a-b62a-bef10b4b8355 · inbound

Reasoning Within the Mind: Dynamic Multimodal Interleaving in Latent Space cites this paper.

Reasoning Within the Mind: Dynamic Multimodal Interleaving in Latent Space GLM-4.5V and GLM-4.1V-Thinking: Towards Versatile Multimodal Reasoning with Scalable Reinforcement Learning

Reference 3

Resolution
verified exact
local_arxiv, observed 2026-05-16T23:03:38.931235Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.

source=pdf_text observed=2026-05-16T23:02:28.588225Z digest=sha256:7abfdd59771c55742971f87f4f1fb42fce6e48598c9a952f0317d2a1dc41ec95

Observation 7f0461e3-46d1-41b9-b3c3-44d0c312626e · inbound

Streaming Video Instruction Tuning cites this paper.

Streaming Video Instruction Tuning GLM-4.5V and GLM-4.1V-Thinking: Towards Versatile Multimodal Reasoning with Scalable Reinforcement Learning

Reference 31

Resolution
verified exact
local_arxiv, observed 2026-05-16T19:48:21.813161Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.

source=pdf_text observed=2026-05-16T19:44:11.032898Z digest=sha256:26a2dc060cee4e0a9a4d75e6916907bcf1f91ce62ee0d71a9e03461ba42fd0df

Observation 18c44e9b-5659-42ca-8b99-1dc690279dc6 · inbound

Molmo2: Open Weights and Data for Vision-Language Models with Video Understanding and Grounding cites this paper.

Molmo2: Open Weights and Data for Vision-Language Models with Video Understanding and Grounding GLM-4.5V and GLM-4.1V-Thinking: Towards Versatile Multimodal Reasoning with Scalable Reinforcement Learning

Reference 137

Resolution
verified exact
local_arxiv, observed 2026-05-16T04:21:29.895201Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.

source=pdf_text observed=2026-05-16T04:21:29.526008Z digest=sha256:b53a155516fcc6f4dc299a330a80e025ed4dd0b1fea0eb1602ec12ac4544921c

Observation c27aff3e-1747-415c-bbb0-2e18a49288c7 · inbound

CodeOCR: On the Effectiveness of Vision Language Models in Code Understanding cites this paper.

CodeOCR: On the Effectiveness of Vision Language Models in Code Understanding GLM-4.5V and GLM-4.1V-Thinking: Towards Versatile Multimodal Reasoning with Scalable Reinforcement Learning

Reference 88

Resolution
verified exact
local_arxiv, observed 2026-05-16T08:32:36.417528Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.

source=pdf_text observed=2026-05-16T08:30:50.984873Z digest=sha256:5bc6b51a5e5102967fa9506b3c2c0036377e51f6a7185a75106636091fcbbf06

Observation 63d80ec6-ecd4-463b-8303-c5004db8a284 · inbound

SpatiaLab: Can Vision-Language Models Perform Spatial Reasoning in the Wild? cites this paper.

SpatiaLab: Can Vision-Language Models Perform Spatial Reasoning in the Wild? GLM-4.5V and GLM-4.1V-Thinking: Towards Versatile Multimodal Reasoning with Scalable Reinforcement Learning

Reference 2

Resolution
metadata mismatch
local_arxiv, observed 2026-05-16T07:57:33.005453Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.

source=pdf_text observed=2026-07-11T11:50:26.030339Z digest=sha256:f92414480222adcb78293f1b585da489bc0105cb2983710d5328227d3ff72ccd

Observation a3b31a65-1a7b-4ee1-96ee-61bdbeb327b4 · inbound

VISTA-Bench: Do Vision-Language Models Really Understand Visualized Text as Well as Pure Text? cites this paper.

VISTA-Bench: Do Vision-Language Models Really Understand Visualized Text as Well as Pure Text? GLM-4.5V and GLM-4.1V-Thinking: Towards Versatile Multimodal Reasoning with Scalable Reinforcement Learning

Reference 6

Resolution
verified exact
local_arxiv, observed 2026-05-21T13:40:12.595718Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.

source=pdf_text observed=2026-05-21T13:36:28.844714Z digest=sha256:134d5499f53bbaf5ff37f3bcd42a3c3cd627130846ee8470ba6a0b9f55e69756

Observation 665bb61f-ac43-4421-b26f-32fae272198d · inbound

SUPERGLASSES: Benchmarking Vision Language Models as Intelligent Agents for AI Smart Glasses cites this paper.

SUPERGLASSES: Benchmarking Vision Language Models as Intelligent Agents for AI Smart Glasses GLM-4.5V and GLM-4.1V-Thinking: Towards Versatile Multimodal Reasoning with Scalable Reinforcement Learning

Reference 18

Resolution
verified exact
local_arxiv, observed 2026-05-15T19:20:16.450466Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.

source=pdf_text observed=2026-05-15T19:18:36.314029Z digest=sha256:2e549da17ae1805d6fcb2cb1d1477a0a502e83633f34cec29335449cdcf023f4

Observation d190f107-bbf6-4d5c-aac3-9387529d2fe7 · inbound

Dual Tuning for Reasoning Efficacy-Driven Data Curation in Multimodal LLM Training cites this paper.

Dual Tuning for Reasoning Efficacy-Driven Data Curation in Multimodal LLM Training GLM-4.5V and GLM-4.1V-Thinking: Towards Versatile Multimodal Reasoning with Scalable Reinforcement Learning

Reference 5

Resolution
verified exact
local_arxiv, observed 2026-05-16T08:12:35.481553Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.

source=pdf_text observed=2026-05-16T08:12:25.063204Z digest=sha256:1a776cbf93dccaae0bcb63658a5bfdd612a890a45f106998df524910326e4139

Observation 4aad42c3-e306-4451-924a-40c9ff6b7a89 · inbound

TIQA: Human-Aligned Perceptual Text Quality Assessment in Generated Images cites this paper.

TIQA: Human-Aligned Perceptual Text Quality Assessment in Generated Images GLM-4.5V and GLM-4.1V-Thinking: Towards Versatile Multimodal Reasoning with Scalable Reinforcement Learning

Reference 51

Resolution
verified exact
local_arxiv, observed 2026-05-15T14:35:55.844251Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.

source=pdf_text observed=2026-05-15T14:32:59.720416Z digest=sha256:2a9765f9165ee2a5e204c9903999250482ed50a526b875d73626359c44464f28

Observation d26eaf49-3a9d-46ac-8400-7e5073358044 · inbound

Reasoning over Video: Evaluating How MLLMs Extract, Integrate, and Reconstruct Spatiotemporal Evidence cites this paper.

Reasoning over Video: Evaluating How MLLMs Extract, Integrate, and Reconstruct Spatiotemporal Evidence GLM-4.5V and GLM-4.1V-Thinking: Towards Versatile Multimodal Reasoning with Scalable Reinforcement Learning

Reference 10

Resolution
verified exact
local_arxiv, observed 2026-05-15T11:29:58.752853Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.

source=pdf_text observed=2026-05-15T11:28:29.772341Z digest=sha256:f722c0df64a949fb1afc118108c15fe7f2f4a8b1df788474794fdc4bc0252a02

Observation 6c5a865c-580b-4e09-bda0-81adaee0d878 · inbound

ForestPrune: High-ratio Visual Token Compression for Video Multimodal Large Language Models via Spatial-Temporal Forest Modeling cites this paper.

ForestPrune: High-ratio Visual Token Compression for Video Multimodal Large Language Models via Spatial-Temporal Forest Modeling GLM-4.5V and GLM-4.1V-Thinking: Towards Versatile Multimodal Reasoning with Scalable Reinforcement Learning

Reference 14

Resolution
verified exact
local_arxiv, observed 2026-05-15T00:58:25.535528Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.

source=pdf_text observed=2026-05-15T00:56:47.841355Z digest=sha256:e0351539c6af698ccc7cd29ff2eb41f2e00ae6e4c57e418275ca0d4c89cf5259

Observation 0829e7d9-7d9b-45d9-ae28-c5f2f7dca782 · inbound

UniICL: Systematizing Unified Multimodal In-context Learning through a Capability-Oriented Taxonomy cites this paper.

UniICL: Systematizing Unified Multimodal In-context Learning through a Capability-Oriented Taxonomy GLM-4.5V and GLM-4.1V-Thinking: Towards Versatile Multimodal Reasoning with Scalable Reinforcement Learning

Reference 13

Resolution
unresolved
no resolver link, observed 2026-07-13T18:42:33.178622Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-13T18:42:33.178622Z digest=sha256:bd8e5bc1a1528e7fe272ffd0a9da341cea8ef09b7569f49e3841222a2f989a89

Observation a6578ab0-b2bd-45d4-8505-a2a07a42b096 · inbound

An Empirical Study of Multi-Agent Collaboration for Automated Research cites this paper.

An Empirical Study of Multi-Agent Collaboration for Automated Research GLM-4.5V and GLM-4.1V-Thinking: Towards Versatile Multimodal Reasoning with Scalable Reinforcement Learning

Reference 2

Resolution
metadata mismatch
local_arxiv, observed 2026-05-13T23:48:27.869428Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.

source=pdf_text observed=2026-05-13T23:43:30.595949Z digest=sha256:85a65344c29b6e2e9722c50b65868364e3001fc2fe30aa17d86589f621c20136

Observation 4888092d-2e6b-4092-9169-c471d3e22492 · inbound

Internalized Reasoning for Long-Context Visual Document Understanding cites this paper.

Internalized Reasoning for Long-Context Visual Document Understanding GLM-4.5V and GLM-4.1V-Thinking: Towards Versatile Multimodal Reasoning with Scalable Reinforcement Learning

Reference 58

Resolution
verified exact
local_arxiv, observed 2026-05-13T23:53:28.008584Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.

source=pdf_text observed=2026-05-13T23:53:19.148407Z digest=sha256:37afd0da9fe99e98dd46c586f473788c597c2a838737d7f764e620e2ba8247d4

Observation 6e043041-0dff-400a-86a3-3e434df7a6ca · inbound

Internalized Reasoning for Long-Context Visual Document Understanding cites this paper.

Internalized Reasoning for Long-Context Visual Document Understanding GLM-4.5V and GLM-4.1V-Thinking: Towards Versatile Multimodal Reasoning with Scalable Reinforcement Learning

Reference 58

Resolution
unresolved
no resolver link, observed 2026-07-13T15:50:49.083652Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-13T15:50:49.083652Z digest=sha256:6ffae28eb35d06695048edfb522c08158d6d779b0572de826247bd28417ef842

Observation 0128bfd9-4c23-4d5d-bca6-c814e0db1630 · inbound

Supply-Chain Poisoning Attacks Against LLM Coding Agent Skill Ecosystems cites this paper.

Supply-Chain Poisoning Attacks Against LLM Coding Agent Skill Ecosystems GLM-4.5V and GLM-4.1V-Thinking: Towards Versatile Multimodal Reasoning with Scalable Reinforcement Learning

Reference 19

Resolution
verified exact
local_arxiv, observed 2026-05-13T19:48:11.205747Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.

source=pdf_text observed=2026-05-13T19:47:45.936608Z digest=sha256:e2498d5a516ee3c567969e5cec73332760b989d7d0802e869381a11c7c8adb4c

Observation a7337034-55da-4371-be1d-acc7ade04578 · inbound

CoME-VL: Scaling Complementary Multi-Encoder Vision-Language Learning cites this paper.

CoME-VL: Scaling Complementary Multi-Encoder Vision-Language Learning GLM-4.5V and GLM-4.1V-Thinking: Towards Versatile Multimodal Reasoning with Scalable Reinforcement Learning

Reference 25

Resolution
verified exact
local_arxiv, observed 2026-05-13T20:33:17.102039Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.

source=pdf_text observed=2026-05-13T20:28:30.864143Z digest=sha256:1ce9055e18601259a6e6164e36b2ea64aa31fc5322f8ea1298c963b2a82cb533

Observation 0e1e5d5d-783f-430d-a9bc-807341bca6bf · inbound

Video-MME-v2: Towards the Next Stage in Benchmarks for Comprehensive Video Understanding cites this paper.

Video-MME-v2: Towards the Next Stage in Benchmarks for Comprehensive Video Understanding GLM-4.5V and GLM-4.1V-Thinking: Towards Versatile Multimodal Reasoning with Scalable Reinforcement Learning

Reference 12

Resolution
verified exact
arxiv_id, observed 2026-05-11T04:48:27.616218Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.

source=pdf_text observed=2026-05-10T18:47:32.778695Z digest=sha256:959ff7d7be1076a002a5c79ddf59aebe40b37798e7a9a59087ed9a21dcdb88f6

Observation f497adda-e158-4e2e-a6e6-34b888c95ba7 · inbound

EpiBench: Benchmarking Multi-turn Research Workflows for Multimodal Agents cites this paper.

EpiBench: Benchmarking Multi-turn Research Workflows for Multimodal Agents GLM-4.5V and GLM-4.1V-Thinking: Towards Versatile Multimodal Reasoning with Scalable Reinforcement Learning

Reference 28

Resolution
verified exact
arxiv_id, observed 2026-05-11T04:48:27.616218Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.

source=pdf_text observed=2026-05-10T18:31:15.525331Z digest=sha256:9cf18468ce6d0c762c4ddf6b084bbfc32d580f6bbd66bc20cb45d7aebf992058

Observation a63fa3d1-5a51-4d3a-b648-d7fd4bb35cef · inbound

DetailVerifyBench: A Benchmark for Dense Hallucination Localization in Long Image Captions cites this paper.

DetailVerifyBench: A Benchmark for Dense Hallucination Localization in Long Image Captions GLM-4.5V and GLM-4.1V-Thinking: Towards Versatile Multimodal Reasoning with Scalable Reinforcement Learning

Reference 44

Resolution
verified exact
arxiv_id, observed 2026-05-11T04:48:27.616218Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.

source=pdf_text observed=2026-05-10T18:39:29.592129Z digest=sha256:a350742c85754015e98380d283d486962a64504a7b8e5ea5ad551d80dfe83c3c

Observation 2a0ae7f4-506c-4060-a33a-c52355aace96 · inbound

MedConclusion: A Benchmark for Biomedical Conclusion Generation from Structured Abstracts cites this paper.

MedConclusion: A Benchmark for Biomedical Conclusion Generation from Structured Abstracts GLM-4.5V and GLM-4.1V-Thinking: Towards Versatile Multimodal Reasoning with Scalable Reinforcement Learning

Reference 24

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T04:48:27.616218Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.

source=pdf_text observed=2026-05-10T18:36:50.613703Z digest=sha256:f9ef895d78ada20b9de926938b5af47749728358f8e6ece73c2e988fa00ead47

Observation 0e71f657-05e2-4345-8aea-d168494f816a · inbound

GameWorld: Towards Standardized and Verifiable Evaluation of Multimodal Game Agents cites this paper.

GameWorld: Towards Standardized and Verifiable Evaluation of Multimodal Game Agents GLM-4.5V and GLM-4.1V-Thinking: Towards Versatile Multimodal Reasoning with Scalable Reinforcement Learning

Reference 33

Resolution
verified exact
local_arxiv, observed 2026-05-11T05:46:07.006010Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.

source=pdf_text observed=2026-05-10T17:57:36.038091Z digest=sha256:7d9c4b7add8eec18808194dec53008d759ac5ae5ee93c4dacef6ec447c489ae0

Observation f2eeac52-ac62-4155-a0c4-3403a56ccf7d · inbound

PokeGym: A Visually-Driven Long-Horizon Benchmark for Vision-Language Models cites this paper.

PokeGym: A Visually-Driven Long-Horizon Benchmark for Vision-Language Models GLM-4.5V and GLM-4.1V-Thinking: Towards Versatile Multimodal Reasoning with Scalable Reinforcement Learning

Reference 70

Resolution
verified exact
local_arxiv, observed 2026-05-11T05:41:00.379965Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.

source=pdf_text observed=2026-05-10T17:59:48.877783Z digest=sha256:29634051f11eefa64a81fb5b0fe7590549eb4dd52d805ad0e58e6a115a7076d1

Observation 37af00b6-d404-447e-ad63-f7c0a63bbc6b · inbound

MolmoWeb: Open Visual Web Agent and Open Data for the Open Web cites this paper.

MolmoWeb: Open Visual Web Agent and Open Data for the Open Web GLM-4.5V and GLM-4.1V-Thinking: Towards Versatile Multimodal Reasoning with Scalable Reinforcement Learning

Reference 28

Resolution
verified exact
local_arxiv, observed 2026-05-11T05:40:58.869456Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.

source=pdf_text observed=2026-05-10T18:00:34.401698Z digest=sha256:174c2c35fb64d93c93b2932ee4dde7fc5b85dbbc5e934e525e187d5e40027794

Observation b5e052d6-d844-45f6-babb-bbb3b42b0d9e · inbound

HM-Bench: A Comprehensive Benchmark for Multimodal Large Language Models in Hyperspectral Remote Sensing cites this paper.

HM-Bench: A Comprehensive Benchmark for Multimodal Large Language Models in Hyperspectral Remote Sensing GLM-4.5V and GLM-4.1V-Thinking: Towards Versatile Multimodal Reasoning with Scalable Reinforcement Learning

Reference 20

Resolution
verified exact
local_arxiv, observed 2026-05-11T05:30:57.716383Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.

source=pdf_text observed=2026-05-10T18:06:49.114269Z digest=sha256:ac25a2316b65430e87343645f678b018cfbac9e1a9f43a245e2bc290f9a93166

Observation 76f18d35-067c-4f3a-8eb9-3d830a62c0f9 · inbound

UIPress: Bringing Optical Token Compression to UI-to-Code Generation cites this paper.

UIPress: Bringing Optical Token Compression to UI-to-Code Generation GLM-4.5V and GLM-4.1V-Thinking: Towards Versatile Multimodal Reasoning with Scalable Reinforcement Learning

Reference 49

Resolution
verified exact
local_arxiv, observed 2026-05-11T07:00:59.478488Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.

source=pdf_text observed=2026-05-10T17:21:32.024105Z digest=sha256:0063eb24aaa61fe89711d05d2e6f9d0e923e976efbee66e922673920c0bc6166

Observation 3c7a35d8-e3ac-4996-b3c9-a2e2a63e463c · inbound

MMRareBench: A Rare-Disease Multimodal and Multi-Image Medical Benchmark cites this paper.

MMRareBench: A Rare-Disease Multimodal and Multi-Image Medical Benchmark GLM-4.5V and GLM-4.1V-Thinking: Towards Versatile Multimodal Reasoning with Scalable Reinforcement Learning

Reference 10

Resolution
metadata mismatch
local_arxiv, observed 2026-05-11T10:21:01.352426Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.

source=pdf_text observed=2026-05-10T15:32:59.205848Z digest=sha256:dea45eb26ea8f197a767f49966b7e9287011f262efb6f6c0011d1a91bf2b0ce2

Observation f43525a7-90e0-483f-9a8c-6d836a29463a · inbound

MMRareBench: A Rare-Disease Multimodal and Multi-Image Medical Benchmark cites this paper.

MMRareBench: A Rare-Disease Multimodal and Multi-Image Medical Benchmark GLM-4.5V and GLM-4.1V-Thinking: Towards Versatile Multimodal Reasoning with Scalable Reinforcement Learning

Reference 10

Resolution
metadata mismatch
local_arxiv, observed 2026-05-14T21:17:58.582638Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.

source=pdf_text observed=2026-05-14T21:17:20.190144Z digest=sha256:3a5b8968c91a67682f4c440a82c1ea0d33ead5e967ae39dd56123a0193b1182d

Observation f81df95c-acf3-4855-8fa6-8977c9e50e80 · inbound

POINTS-Long: Adaptive Dual-Mode Visual Reasoning in MLLMs cites this paper.

POINTS-Long: Adaptive Dual-Mode Visual Reasoning in MLLMs GLM-4.5V and GLM-4.1V-Thinking: Towards Versatile Multimodal Reasoning with Scalable Reinforcement Learning

Reference 81

Resolution
verified exact
local_arxiv, observed 2026-05-11T10:41:04.354487Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.

source=pdf_text observed=2026-05-10T15:23:08.671342Z digest=sha256:a29ce5f75f7767e517d88ba4fa1649fad7f2c6a6a64e7f9666dbadba9ff488b3

Observation c1666d75-9171-4363-8d40-8051121b4efa · inbound

Towards Robust Real-World Spreadsheet Understanding with Multi-Agent Multi-Format Reasoning cites this paper.

Towards Robust Real-World Spreadsheet Understanding with Multi-Agent Multi-Format Reasoning GLM-4.5V and GLM-4.1V-Thinking: Towards Versatile Multimodal Reasoning with Scalable Reinforcement Learning

Reference 8

Resolution
verified exact
arxiv_id, observed 2026-05-11T04:48:27.616218Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.

source=arxiv_source observed=2026-05-10T15:32:54.053527Z digest=sha256:d24ec4c6744d4c70cf5dc2140f2440060545ed382c5f97166f3526955d40b2d6

Observation 22751ca3-1fa3-42b4-86ed-72ee9de514a1 · inbound

Why and When Visual Token Pruning Fails? A Study on Relevant Visual Information Shift in MLLMs Decoding cites this paper.

Why and When Visual Token Pruning Fails? A Study on Relevant Visual Information Shift in MLLMs Decoding GLM-4.5V and GLM-4.1V-Thinking: Towards Versatile Multimodal Reasoning with Scalable Reinforcement Learning

Reference 40

Resolution
verified exact
local_arxiv, observed 2026-05-11T11:31:01.391919Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.

source=pdf_text observed=2026-05-10T14:50:37.022338Z digest=sha256:fdc95f842aa77df314a1e699e40e4e0cca0d4c4f11306f7aaf6d7bcbe09460d3

Observation 0400511a-8c66-4178-8e4d-f5134b8ad98e · inbound

Every Picture Tells a Dangerous Story: Memory-Augmented Multi-Agent Jailbreak Attacks on VLMs cites this paper.

Every Picture Tells a Dangerous Story: Memory-Augmented Multi-Agent Jailbreak Attacks on VLMs GLM-4.5V and GLM-4.1V-Thinking: Towards Versatile Multimodal Reasoning with Scalable Reinforcement Learning

Reference 48

Resolution
verified exact
local_arxiv, observed 2026-05-11T11:11:05.467585Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.

source=pdf_text observed=2026-05-10T15:06:43.140580Z digest=sha256:2794a44078fc128ccfcbff382a4b606d76bdcb4ceef7308d208304d5f7046404

Observation 6aa484e7-3653-45b9-a2fb-68000744cbed · inbound

DocSeeker: Structured Visual Reasoning with Evidence Grounding for Long Document Understanding cites this paper.

DocSeeker: Structured Visual Reasoning with Evidence Grounding for Long Document Understanding GLM-4.5V and GLM-4.1V-Thinking: Towards Versatile Multimodal Reasoning with Scalable Reinforcement Learning

Reference 12

Resolution
verified exact
local_arxiv, observed 2026-05-11T09:05:58.510332Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.

source=pdf_text observed=2026-05-10T16:16:58.889065Z digest=sha256:3721bb09432fdb0a84d6478efe7f90b6b5949ae316ac2c9045dc4a5c124b0b91

Observation 93677088-9fc0-4b52-9cd7-417245b290a3 · inbound

DocSeeker: Structured Visual Reasoning with Evidence Grounding for Long Document Understanding cites this paper.

DocSeeker: Structured Visual Reasoning with Evidence Grounding for Long Document Understanding GLM-4.5V and GLM-4.1V-Thinking: Towards Versatile Multimodal Reasoning with Scalable Reinforcement Learning

Reference 12

Resolution
verified exact
local_arxiv, observed 2026-05-12T06:26:24.479026Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.

source=pdf_text observed=2026-05-12T04:17:55.318813Z digest=sha256:ba7dd135f8cec6dd9e4b8683102c03c5d4985ad787b73ae665c0521a446a2fe4

Observation 888eb7aa-7eaf-4144-b8a9-766de0f543af · inbound

Grasp in Gaussians: Fast Monocular Reconstruction of Dynamic Hand-Object Interactions cites this paper.

Grasp in Gaussians: Fast Monocular Reconstruction of Dynamic Hand-Object Interactions GLM-4.5V and GLM-4.1V-Thinking: Towards Versatile Multimodal Reasoning with Scalable Reinforcement Learning

Reference 51

Resolution
metadata mismatch
local_arxiv, observed 2026-05-11T11:21:00.498276Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.

source=pdf_text observed=2026-05-10T15:01:52.354836Z digest=sha256:4fa7099d2859d7cc69a6b6fb5d9d14e242eadfe2d6810c145858e5cbe23bbf37

Observation 1b09a0da-7c9d-4814-9b01-bfd44cc74e5e · inbound

RiskWebWorld: A Realistic Interactive Benchmark for GUI Agents in E-commerce Risk Management cites this paper.

RiskWebWorld: A Realistic Interactive Benchmark for GUI Agents in E-commerce Risk Management GLM-4.5V and GLM-4.1V-Thinking: Towards Versatile Multimodal Reasoning with Scalable Reinforcement Learning

Reference 16

Resolution
verified exact
arxiv_id, observed 2026-05-11T04:48:27.616218Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.

source=pdf_text observed=2026-05-10T13:40:55.540869Z digest=sha256:bf2851ff21048b64925c070a7a190f5e1e52d30a1bfd074797659f594d9ae16f

Observation 9b7a13d4-c3cb-4a69-887b-cdbc26a71b74 · inbound

MirrorBench: Evaluating Self-centric Intelligence in MLLMs by Introducing a Mirror cites this paper.

MirrorBench: Evaluating Self-centric Intelligence in MLLMs by Introducing a Mirror GLM-4.5V and GLM-4.1V-Thinking: Towards Versatile Multimodal Reasoning with Scalable Reinforcement Learning

Reference 34

Resolution
verified exact
arxiv_id, observed 2026-05-11T04:48:27.616218Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.

source=pdf_text observed=2026-05-10T10:53:48.374637Z digest=sha256:208eb93df4da82d7872e7abac57911bc838dc2f279bc276cefcbcba1c2edcb9b

Observation 25c3b223-320a-4250-99b3-58a83036e7ce · inbound

OASIS: On-Demand Hierarchical Event Memory for Streaming Video Reasoning cites this paper.

OASIS: On-Demand Hierarchical Event Memory for Streaming Video Reasoning GLM-4.5V and GLM-4.1V-Thinking: Towards Versatile Multimodal Reasoning with Scalable Reinforcement Learning

Reference 18

Resolution
verified exact
arxiv_id, observed 2026-05-11T04:48:27.616218Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.

source=pdf_text observed=2026-05-10T06:51:52.861981Z digest=sha256:df7a8ada9b252fc2d3ceda8ae795a59cf6b3f2ce06f9ee1c96b7613c3114ccee

Observation 9dbacffd-9dd8-4239-bafc-960e4dff3872 · inbound

UCCL-Zip: Lossless Compression Supercharged GPU Communication cites this paper.

UCCL-Zip: Lossless Compression Supercharged GPU Communication GLM-4.5V and GLM-4.1V-Thinking: Towards Versatile Multimodal Reasoning with Scalable Reinforcement Learning

Reference 52

Resolution
verified exact
arxiv_id, observed 2026-05-11T04:48:27.616218Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.

source=pdf_text observed=2026-05-10T06:36:43.101810Z digest=sha256:355dd42c14a6bc47cdaa70db7b827e5b891636e58e119170aaed1e30f780d329

Observation 62813321-8680-4086-b857-6f965d1a5945 · inbound

SpatialImaginer: Towards Adaptive Visual Imagination for Spatial Reasoning cites this paper.

SpatialImaginer: Towards Adaptive Visual Imagination for Spatial Reasoning GLM-4.5V and GLM-4.1V-Thinking: Towards Versatile Multimodal Reasoning with Scalable Reinforcement Learning

Reference 20

Resolution
verified exact
arxiv_id, observed 2026-05-11T04:48:27.616218Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.

source=pdf_text observed=2026-05-10T06:31:30.778309Z digest=sha256:b29345bf55c4e547d7f171dd47f34f8ca91411bb8d11d3e6cc0ea13b198a7314

Observation 92382599-35b1-4db0-a685-3c2fefa791a5 · inbound

Back into Plato's Cave: Examining Cross-modal Representational Convergence at Scale cites this paper.

Back into Plato's Cave: Examining Cross-modal Representational Convergence at Scale GLM-4.5V and GLM-4.1V-Thinking: Towards Versatile Multimodal Reasoning with Scalable Reinforcement Learning

Reference 35

Resolution
verified exact
arxiv_id, observed 2026-05-11T04:48:27.616218Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.

source=pdf_text observed=2026-05-10T05:00:31.717115Z digest=sha256:a5c02f1e8bdf6b2c246980b46a7e8fcf41cd05fe1ef5fa6d3dabfdb7dc125298

Observation 2e2a2385-da03-429e-89b7-719fe78fadb0 · inbound

SAMoRA: Semantic-Aware Mixture of LoRA Experts for Task-Adaptive Learning cites this paper.

SAMoRA: Semantic-Aware Mixture of LoRA Experts for Task-Adaptive Learning GLM-4.5V and GLM-4.1V-Thinking: Towards Versatile Multimodal Reasoning with Scalable Reinforcement Learning

Reference 60

Resolution
verified exact
arxiv_id, observed 2026-05-11T04:48:27.616218Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.

source=arxiv_source observed=2026-05-10T02:26:55.896058Z digest=sha256:8a00662b180c5d57781a98f15a3ccfbb6f8692abc342b33f9824eaa2a2d6ba8a

Observation 41e74320-884e-4b83-a233-b4ed61a4eb34 · inbound

HyLaR: Hybrid Latent Reasoning with Decoupled Policy Optimization cites this paper.

HyLaR: Hybrid Latent Reasoning with Decoupled Policy Optimization GLM-4.5V and GLM-4.1V-Thinking: Towards Versatile Multimodal Reasoning with Scalable Reinforcement Learning

Reference 13

Resolution
metadata mismatch
local_arxiv, observed 2026-05-11T13:41:04.095036Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.

source=pdf_text observed=2026-05-10T01:17:33.668451Z digest=sha256:edd503fe6d559293b89e7b4de7dce6dc890937f4e3039a040490329c0dc1487d

Observation 54642419-d86a-4860-8dc9-909ce3a9fa88 · inbound

HyLaR: Hybrid Latent Reasoning with Decoupled Policy Optimization cites this paper.

HyLaR: Hybrid Latent Reasoning with Decoupled Policy Optimization GLM-4.5V and GLM-4.1V-Thinking: Towards Versatile Multimodal Reasoning with Scalable Reinforcement Learning

Reference 13

Resolution
metadata mismatch
local_arxiv, observed 2026-07-05T04:30:40.733541Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.

source=pdf_text observed=2026-07-05T04:29:52.480873Z digest=sha256:061060edcc1892bf9c4a10f97246f59f23ddee744f99dc4e1c752ddef66ff7f0

Observation 35d8cbce-a93e-45bd-a3d9-81a0247cd289 · inbound

X-PCR: A Benchmark for Cross-modality Progressive Clinical Reasoning in Ophthalmic Diagnosis cites this paper.

X-PCR: A Benchmark for Cross-modality Progressive Clinical Reasoning in Ophthalmic Diagnosis GLM-4.5V and GLM-4.1V-Thinking: Towards Versatile Multimodal Reasoning with Scalable Reinforcement Learning

Reference 13

Resolution
verified exact
arxiv_id, observed 2026-05-11T04:48:27.616218Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.

source=pdf_text observed=2026-05-10T01:04:07.511600Z digest=sha256:85f292eb42c98f8069af9a0b4924d1f84588be480a94022fa6fdacc960a1eb5f

Observation a1f24032-cb80-4956-9d9a-04c72cec029a · inbound

Towards Temporal Compositional Reasoning in Long-Form Sports Videos cites this paper.

Towards Temporal Compositional Reasoning in Long-Form Sports Videos GLM-4.5V and GLM-4.1V-Thinking: Towards Versatile Multimodal Reasoning with Scalable Reinforcement Learning

Reference 30

Resolution
metadata mismatch
local_arxiv, observed 2026-05-11T19:06:08.908919Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.

source=pdf_text observed=2026-05-08T12:45:50.422679Z digest=sha256:8db8bdb66b1a7702754c98cd0b5c0f388a85cfa4fff1901a31025596cf6c6743

Observation 49742b1b-818c-46b5-b4fe-76bfbec4b988 · inbound

Towards Temporal Compositional Reasoning in Long-Form Sports Videos cites this paper.

Towards Temporal Compositional Reasoning in Long-Form Sports Videos GLM-4.5V and GLM-4.1V-Thinking: Towards Versatile Multimodal Reasoning with Scalable Reinforcement Learning

Reference 30

Resolution
unresolved
no resolver link, observed 2026-07-14T19:27:29.843866Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T19:27:29.843866Z digest=sha256:5bfee3b94fbf641a02b2bdaf0f8ccdc6b40af5dcbad8e7184190f17e01deb54c

Observation 615ba1d5-a79a-4ff4-a91a-cc32a2e02b40 · inbound

SOLAR-RL: Semi-Online Long-horizon Assignment Reinforcement Learning cites this paper.

SOLAR-RL: Semi-Online Long-horizon Assignment Reinforcement Learning GLM-4.5V and GLM-4.1V-Thinking: Towards Versatile Multimodal Reasoning with Scalable Reinforcement Learning

Reference 6

Resolution
verified exact
local_arxiv, observed 2026-05-11T19:21:08.864116Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.

source=arxiv_source observed=2026-05-08T12:09:24.371878Z digest=sha256:400b6f59951cd5ecb0237c4f14718defe419af1e8146cd9d016aa1f0ad833c35

Observation 07440cff-76da-4452-89a5-20bd170dda9e · inbound

AeSlides: Incentivizing Aesthetic Layout in LLM-Based Slide Generation via Verifiable Rewards cites this paper.

AeSlides: Incentivizing Aesthetic Layout in LLM-Based Slide Generation via Verifiable Rewards GLM-4.5V and GLM-4.1V-Thinking: Towards Versatile Multimodal Reasoning with Scalable Reinforcement Learning

Reference 7

Resolution
verified exact
local_arxiv, observed 2026-05-11T13:06:06.001569Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.

source=pdf_text observed=2026-05-10T02:18:05.770211Z digest=sha256:5ffab156baa6dcaed369a5a50ec12140e0ea8b6bc3d4e5c430029cb823e23f8e

Observation 44678f3e-8fe7-4ef5-844f-b3b260aeff71 · inbound

OS-SPEAR: A Toolkit for the Safety, Performance,Efficiency, and Robustness Analysis of OS Agents cites this paper.

OS-SPEAR: A Toolkit for the Safety, Performance,Efficiency, and Robustness Analysis of OS Agents GLM-4.5V and GLM-4.1V-Thinking: Towards Versatile Multimodal Reasoning with Scalable Reinforcement Learning

Reference 94

Resolution
verified exact
local_arxiv, observed 2026-05-11T21:56:11.922010Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.

source=pdf_text observed=2026-05-08T03:51:54.310805Z digest=sha256:1f642b45709029de9d51b9746f77e496020776077fd3f37cdf4bb9242afb1174

Observation 8cf4b577-08a0-4887-b9c4-52e91c47ca99 · inbound

MAIC-UI: Making Interactive Courseware with Generative UI cites this paper.

MAIC-UI: Making Interactive Courseware with Generative UI GLM-4.5V and GLM-4.1V-Thinking: Towards Versatile Multimodal Reasoning with Scalable Reinforcement Learning

Reference 21

Resolution
verified exact
local_arxiv, observed 2026-05-11T23:41:16.190174Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.

source=pdf_text observed=2026-05-07T16:31:14.775887Z digest=sha256:01d9ba013f77531569e814b5261e6f01734684fd3f2fd18b39d4c3fd71201437

Observation 2db88b3a-f500-4da6-8045-80dbbba9d601 · inbound

Purifying Multimodal Retrieval: Fragment-Level Evidence Selection for RAG cites this paper.

Purifying Multimodal Retrieval: Fragment-Level Evidence Selection for RAG GLM-4.5V and GLM-4.1V-Thinking: Towards Versatile Multimodal Reasoning with Scalable Reinforcement Learning

Reference 12

Resolution
verified exact
local_arxiv, observed 2026-05-12T10:11:28.117403Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.

source=pdf_text observed=2026-05-07T07:19:44.125479Z digest=sha256:ba94f24ee0523e99ea8d980c0a9d56bc6db492311ce83955c2f7cad987892fc1

Observation 267d1350-c38a-4ca0-bf6d-2cc1126112fa · inbound

Persistent Visual Memory: Sustaining Perception for Deep Generation in LVLMs cites this paper.

Persistent Visual Memory: Sustaining Perception for Deep Generation in LVLMs GLM-4.5V and GLM-4.1V-Thinking: Towards Versatile Multimodal Reasoning with Scalable Reinforcement Learning

Reference 33

Resolution
verified exact
local_arxiv, observed 2026-05-11T16:01:22.921669Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.

source=pdf_text observed=2026-05-09T18:53:06.494640Z digest=sha256:df8de7be5186cabc959b2fef07eeb431696c26b434529f7b0752b8724435aa30

Observation b62ed2c1-c01a-4d14-b46a-d69bdc82a006 · inbound

Persistent Visual Memory: Sustaining Perception for Deep Generation in LVLMs cites this paper.

Persistent Visual Memory: Sustaining Perception for Deep Generation in LVLMs GLM-4.5V and GLM-4.1V-Thinking: Towards Versatile Multimodal Reasoning with Scalable Reinforcement Learning

Reference 33

Resolution
verified exact
arxiv_id, observed 2026-05-11T04:48:27.616218Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.

source=pdf_text observed=2026-05-11T01:49:15.136031Z digest=sha256:5e152c1350f5794256826de83d3cd9536b22ef8296ecaddaedafe8225824d78d

Observation 8e8cba20-9306-4351-b01a-b497090f2b0f · inbound

VL-SAM-v3: Memory-Guided Visual Priors for Open-World Object Detection cites this paper.

VL-SAM-v3: Memory-Guided Visual Priors for Open-World Object Detection GLM-4.5V and GLM-4.1V-Thinking: Towards Versatile Multimodal Reasoning with Scalable Reinforcement Learning

Reference 15

Resolution
verified exact
local_arxiv, observed 2026-05-12T11:01:30.251232Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.

source=pdf_text observed=2026-05-08T01:24:56.236650Z digest=sha256:357b292231d4187e0cd0163f5f74d8afa340b0e74dcf6e9828376d0e970187fb

Observation 935adb88-c659-4040-ab92-d51f559f1556 · inbound

VL-SAM-v3: Memory-Guided Visual Priors for Open-World Object Detection cites this paper.

VL-SAM-v3: Memory-Guided Visual Priors for Open-World Object Detection GLM-4.5V and GLM-4.1V-Thinking: Towards Versatile Multimodal Reasoning with Scalable Reinforcement Learning

Reference 15

Resolution
verified exact
arxiv_id, observed 2026-05-11T04:48:27.616218Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.

source=pdf_text observed=2026-05-08T19:14:12.092478Z digest=sha256:f7c0b5c852c142f960d96e399f412c3a72b9ac807c4f422d39a849b092b07292

Observation eff6ba23-677e-498c-ba15-b2d125256af3 · inbound

VL-SAM-v3: Memory-Guided Visual Priors for Open-World Object Detection cites this paper.

VL-SAM-v3: Memory-Guided Visual Priors for Open-World Object Detection GLM-4.5V and GLM-4.1V-Thinking: Towards Versatile Multimodal Reasoning with Scalable Reinforcement Learning

Reference 15

Resolution
verified exact
local_arxiv, observed 2026-05-12T07:51:48.384089Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.

source=pdf_text observed=2026-05-12T01:33:55.317107Z digest=sha256:5bdf27892ec4d04c82c8cea37624a73608fede2794a9d1ebaafda29a3e2760da

Observation 3ab22951-39f3-4335-addb-5a00006360a5 · inbound

MolRecBench-Wild: A Real-World Benchmark for Optical Chemical Structure Recognition cites this paper.

MolRecBench-Wild: A Real-World Benchmark for Optical Chemical Structure Recognition GLM-4.5V and GLM-4.1V-Thinking: Towards Versatile Multimodal Reasoning with Scalable Reinforcement Learning

Reference 9

Resolution
verified exact
local_arxiv, observed 2026-05-11T19:41:08.283888Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.

source=pdf_text observed=2026-05-08T11:20:09.769711Z digest=sha256:19c8e12bfa85c7df6e6bf09921c8ce32ba334c838350af7e41de6c0d92662758

Observation 4832bb9d-8ee8-4079-abdf-3eff57696c16 · inbound

RobotEQ: Transitioning from Passive Intelligence to Active Intelligence in Embodied AI cites this paper.

RobotEQ: Transitioning from Passive Intelligence to Active Intelligence in Embodied AI GLM-4.5V and GLM-4.1V-Thinking: Towards Versatile Multimodal Reasoning with Scalable Reinforcement Learning

Reference 22

Resolution
verified exact
local_arxiv, observed 2026-05-11T20:26:10.086438Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.

source=pdf_text observed=2026-05-08T09:12:00.835585Z digest=sha256:690ab048bc2cd4fcda76bd655942cbb2b68a1a15ca32fbf6460104bf2d0f0de4

Observation b76a2b1f-a841-41fb-be85-cc10b149928d · inbound

RobotEQ: Transitioning from Passive Intelligence to Active Intelligence in Embodied AI cites this paper.

RobotEQ: Transitioning from Passive Intelligence to Active Intelligence in Embodied AI GLM-4.5V and GLM-4.1V-Thinking: Towards Versatile Multimodal Reasoning with Scalable Reinforcement Learning

Reference 22

Resolution
verified exact
local_arxiv, observed 2026-06-30T23:35:07.744441Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.

source=pdf_text observed=2026-06-30T23:27:32.526304Z digest=sha256:7df1aba7750040843f5d746ffbece858e90659a5c1995d58a5f442023636d851

Observation 0058aa81-88f8-41ad-ad39-6b1690f440b9 · inbound

Hard to Read, Easy to Jailbreak: How Visual Degradation Bypasses MLLM Safety Alignment cites this paper.

Hard to Read, Easy to Jailbreak: How Visual Degradation Bypasses MLLM Safety Alignment GLM-4.5V and GLM-4.1V-Thinking: Towards Versatile Multimodal Reasoning with Scalable Reinforcement Learning

Reference 55

Resolution
verified exact
arxiv_id, observed 2026-05-11T04:48:27.616218Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.

source=arxiv_source observed=2026-05-11T01:53:46.652759Z digest=sha256:ef1e6c8f0ae6756046689d0d38f5d9dd2719526ecc3779505566d5fa1af3eff9

Observation 42b15f8b-58c4-4e55-b105-66d20b1088dd · inbound

SphereVAD: Training-Free Video Anomaly Detection via Geodesic Inference on the Unit Hypersphere cites this paper.

SphereVAD: Training-Free Video Anomaly Detection via Geodesic Inference on the Unit Hypersphere GLM-4.5V and GLM-4.1V-Thinking: Towards Versatile Multimodal Reasoning with Scalable Reinforcement Learning

Reference 13

Resolution
verified exact
arxiv_id, observed 2026-05-11T04:48:27.616218Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.

source=pdf_text observed=2026-05-11T02:56:12.267144Z digest=sha256:d39f1ba5634da67725244f0bdc71809b95ba269704b6764c2125268aba47b07d

Observation e731b2fb-a456-4397-9911-00e3eabb5758 · inbound

VT-Bench: A Unified Benchmark for Visual-Tabular Multi-Modal Learning cites this paper.

VT-Bench: A Unified Benchmark for Visual-Tabular Multi-Modal Learning GLM-4.5V and GLM-4.1V-Thinking: Towards Versatile Multimodal Reasoning with Scalable Reinforcement Learning

Reference 43

Resolution
metadata mismatch
local_arxiv, observed 2026-05-12T07:56:27.065120Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.

source=arxiv_source observed=2026-05-12T01:31:52.536354Z digest=sha256:0ca372193c3ea29c136cbc36b672b81a2bb0ee8e63b02bff52d74000b14fa174

Observation e06f919c-1aa0-4eb6-9310-18ce54a1f0e0 · inbound

VT-Bench: A Unified Benchmark for Visual-Tabular Multi-Modal Learning cites this paper.

VT-Bench: A Unified Benchmark for Visual-Tabular Multi-Modal Learning GLM-4.5V and GLM-4.1V-Thinking: Towards Versatile Multimodal Reasoning with Scalable Reinforcement Learning

Reference 43

Resolution
metadata mismatch
local_arxiv, observed 2026-05-21T00:33:52.781109Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.

source=arxiv_source observed=2026-05-21T00:31:07.580063Z digest=sha256:85a423c7aa37a170220acc049d13db35452f31d8cb94b016e7b7358c4498d113

Observation 04e26437-8037-467f-97b7-8c292930d701 · inbound

VT-Bench: A Unified Benchmark for Visual-Tabular Multi-Modal Learning cites this paper.

VT-Bench: A Unified Benchmark for Visual-Tabular Multi-Modal Learning GLM-4.5V and GLM-4.1V-Thinking: Towards Versatile Multimodal Reasoning with Scalable Reinforcement Learning

Reference 1

Resolution
metadata mismatch
local_arxiv, observed 2026-07-01T00:25:09.540399Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.

source=pdf_text observed=2026-07-01T00:20:19.699522Z digest=sha256:25a2edaf7d24dea99d5cafb4cb8b10681a3d23e98a80504f3e68f980303a31db

Observation bcd132f7-d687-4feb-838c-9fa298d9ccf5 · inbound

CustomerSim: Benchmarking and Aligning Multimodal Language Models as Retail User Simulators cites this paper.

CustomerSim: Benchmarking and Aligning Multimodal Language Models as Retail User Simulators GLM-4.5V and GLM-4.1V-Thinking: Towards Versatile Multimodal Reasoning with Scalable Reinforcement Learning

Reference 39

Resolution
verified exact
local_arxiv, observed 2026-05-12T08:41:24.031189Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.

source=pdf_text observed=2026-05-12T00:51:57.796883Z digest=sha256:cc7b4d253ddc7fa3391417b54c8c0eaa5ecec68b459238442db9bb5ee613f4c3

Observation bc54c21c-69ce-4ce1-8ed0-42213428f5f1 · inbound

NICE FACT: Diagnosing and Calibrating VLMs in Quantitative Reasoning for Kinematic Physics cites this paper.

NICE FACT: Diagnosing and Calibrating VLMs in Quantitative Reasoning for Kinematic Physics GLM-4.5V and GLM-4.1V-Thinking: Towards Versatile Multimodal Reasoning with Scalable Reinforcement Learning

Reference 13

Resolution
verified exact
local_arxiv, observed 2026-05-12T07:41:31.935060Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.

source=pdf_text observed=2026-05-12T02:23:04.253007Z digest=sha256:323d0cfdf25cf52a594ae54062d143cdc34fbfb2445ee9e01e118dfa16ad3f7a

Observation 65afc766-0efd-4078-8373-53f248a51301 · inbound

UniShield: Unified Face Attack Detection via KG-Informed Multimodal Reasoning cites this paper.

UniShield: Unified Face Attack Detection via KG-Informed Multimodal Reasoning GLM-4.5V and GLM-4.1V-Thinking: Towards Versatile Multimodal Reasoning with Scalable Reinforcement Learning

Reference 23

Resolution
verified exact
local_arxiv, observed 2026-05-12T08:26:24.389514Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.

source=pdf_text observed=2026-05-12T01:10:21.661074Z digest=sha256:2626c7606f1e903567aba28697969f5e14e48da2ae0a3bfb3b54fdd7889924e2

Observation 72fcc410-93e0-4801-9123-1a10dee9cd49 · inbound

Mem-W: Latent Memory-Native GUI Agents cites this paper.

Mem-W: Latent Memory-Native GUI Agents GLM-4.5V and GLM-4.1V-Thinking: Towards Versatile Multimodal Reasoning with Scalable Reinforcement Learning

Reference 3

Resolution
metadata mismatch
local_arxiv, observed 2026-05-12T06:16:29.126529Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.

source=pdf_text observed=2026-05-12T04:25:01.722487Z digest=sha256:514e27b127c142bae473e8e3a2eaa1674afad225f1451454b738a368b772968a

Observation e20c92e1-5d45-4eb3-b4bb-c23b1756c8e4 · inbound

Chronicles-OCR: A Cross-Temporal Perception Benchmark for the Evolutionary Trajectory of Chinese Characters cites this paper.

Chronicles-OCR: A Cross-Temporal Perception Benchmark for the Evolutionary Trajectory of Chinese Characters GLM-4.5V and GLM-4.1V-Thinking: Towards Versatile Multimodal Reasoning with Scalable Reinforcement Learning

Reference 76

Resolution
verified exact
local_arxiv, observed 2026-05-13T05:52:22.578907Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.

source=pdf_text observed=2026-05-13T05:48:44.584051Z digest=sha256:e03cfcdceefa16b81e2fb6d1822cbcbe4a181e408a93cbdab882800609de78cc

Observation 910d930d-9505-4e90-aa85-34bfabe303fc · inbound

SenseNova-U1: Unifying Multimodal Understanding and Generation with NEO-unify Architecture cites this paper.

SenseNova-U1: Unifying Multimodal Understanding and Generation with NEO-unify Architecture GLM-4.5V and GLM-4.1V-Thinking: Towards Versatile Multimodal Reasoning with Scalable Reinforcement Learning

Reference 52

Resolution
verified exact
local_arxiv, observed 2026-05-13T05:17:18.674337Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.

source=pdf_text observed=2026-05-13T05:12:37.339084Z digest=sha256:9f696331ed6e6854822bb90d8d61bbf6a7fd9c35811977554b96ec5666574646

Observation 5b0b0f08-1861-408e-a1d8-8e538c94fb45 · inbound

Training Long-Context Vision-Language Models Effectively with Generalization Beyond 128K Context cites this paper.

Training Long-Context Vision-Language Models Effectively with Generalization Beyond 128K Context GLM-4.5V and GLM-4.1V-Thinking: Towards Versatile Multimodal Reasoning with Scalable Reinforcement Learning

Reference 14

Resolution
verified exact
local_arxiv, observed 2026-05-14T19:17:50.140441Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.

source=pdf_text observed=2026-05-14T19:16:07.851098Z digest=sha256:0d8a5ce1705d5636a0710f6e3a252d9a5490d09ed805d19fb610b1e502fb2014

Observation 89f2234c-ecd8-4e5c-96c3-daba9360a9d9 · inbound

To See is Not to Learn: Protecting Multimodal Data from Unauthorized Fine-Tuning of Large Vision-Language Model cites this paper.

To See is Not to Learn: Protecting Multimodal Data from Unauthorized Fine-Tuning of Large Vision-Language Model GLM-4.5V and GLM-4.1V-Thinking: Towards Versatile Multimodal Reasoning with Scalable Reinforcement Learning

Reference 19

Resolution
metadata mismatch
local_arxiv, observed 2026-05-15T02:39:40.896473Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.

source=arxiv_source observed=2026-05-15T02:38:37.358485Z digest=sha256:2622274f46e7c79e90d3dd5c94900fccac224afcd1e11153280fc212f977ec9b

Observation 6ddb6e5f-7323-4d57-839e-8cf2ef724733 · inbound

Known By Their Actions: Fingerprinting LLM Browser Agents via UI Traces cites this paper.

Known By Their Actions: Fingerprinting LLM Browser Agents via UI Traces GLM-4.5V and GLM-4.1V-Thinking: Towards Versatile Multimodal Reasoning with Scalable Reinforcement Learning

Reference 38

Resolution
verified exact
local_arxiv, observed 2026-06-30T20:35:02.262380Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.

source=pdf_text observed=2026-06-30T20:34:38.525776Z digest=sha256:bfd8c776f062cdf2678e3fa2f2b3c0d94e00ad34425b4557fdad1441ea28d454

Observation ca8b9b54-550c-4c75-a874-9523bd6e91e3 · inbound

MemLens: Benchmarking Multimodal Long-Term Memory in Large Vision-Language Models cites this paper.

MemLens: Benchmarking Multimodal Long-Term Memory in Large Vision-Language Models GLM-4.5V and GLM-4.1V-Thinking: Towards Versatile Multimodal Reasoning with Scalable Reinforcement Learning

Reference 80

Resolution
verified exact
local_arxiv, observed 2026-06-30T21:05:04.269776Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.

source=pdf_text observed=2026-06-30T21:00:25.664841Z digest=sha256:eca39cce29f8160056008342ceb09816a1a510b0e789c0eca9dd0ca5ca2c605d

Observation 2cf03042-27aa-4fff-b330-dd7592fabdac · inbound

PAGER: Bridging the Semantic-Execution Gap in Point-Precise Geometric GUI Control cites this paper.

PAGER: Bridging the Semantic-Execution Gap in Point-Precise Geometric GUI Control GLM-4.5V and GLM-4.1V-Thinking: Towards Versatile Multimodal Reasoning with Scalable Reinforcement Learning

Reference 14

Resolution
verified exact
local_arxiv, observed 2026-05-20T18:48:53.318235Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.

source=pdf_text observed=2026-05-20T18:46:20.417039Z digest=sha256:0216eeafa9c0069d5fcf6ce7f3dada01870f4ba23dba0daba3f11711d1814c98

Observation 1c71f3f7-b304-4afc-9ad5-02879aaa539f · inbound

HEED: Density-Weighted Residual Alignment for Hybrid Vision-Language Model Distillation cites this paper.

HEED: Density-Weighted Residual Alignment for Hybrid Vision-Language Model Distillation GLM-4.5V and GLM-4.1V-Thinking: Towards Versatile Multimodal Reasoning with Scalable Reinforcement Learning

Reference 37

Resolution
verified exact
local_arxiv, observed 2026-05-20T15:38:26.639128Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.

source=pdf_text observed=2026-05-20T15:37:09.899037Z digest=sha256:739ec108bc7e0ef2aa8a039d77402682b26f48abe042cd4b1940b7b399ae827c

Observation e9301826-9a2d-4d83-9bfb-b343eefbff45 · inbound

Artificial Intolerance: Stigmatizing Language in Clinical Documentation Skews Large Language Model Decision-Making cites this paper.

Artificial Intolerance: Stigmatizing Language in Clinical Documentation Skews Large Language Model Decision-Making GLM-4.5V and GLM-4.1V-Thinking: Towards Versatile Multimodal Reasoning with Scalable Reinforcement Learning

Reference 100

Resolution
metadata mismatch
local_arxiv, observed 2026-05-20T14:53:23.376726Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.

source=arxiv_source observed=2026-05-20T14:48:55.203993Z digest=sha256:5adefd241c282ef14e28cf5564864bbcc1d89f24daf76f3ceb0749d403339fff

Observation 4cfa3ace-8fdf-4f21-8019-054710033a95 · inbound

What's Holding Back Latent Visual Reasoning? cites this paper.

What's Holding Back Latent Visual Reasoning? GLM-4.5V and GLM-4.1V-Thinking: Towards Versatile Multimodal Reasoning with Scalable Reinforcement Learning

Reference 16

Resolution
verified exact
local_arxiv, observed 2026-05-20T12:03:15.440872Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.

source=pdf_text observed=2026-05-20T11:59:14.134917Z digest=sha256:974ec5ccf36bd66ada597bd8ac7e23fae6bd7ae3f090bb671cf66f1df0bc1a91

Observation ea86a291-4383-42e3-97ca-ed5ac71a190c · inbound

Vision-OPD: Learning to See Fine Details for Multimodal LLMs via On-Policy Self-Distillation cites this paper.

Vision-OPD: Learning to See Fine Details for Multimodal LLMs via On-Policy Self-Distillation GLM-4.5V and GLM-4.1V-Thinking: Towards Versatile Multimodal Reasoning with Scalable Reinforcement Learning

Reference 13

Resolution
verified exact
local_arxiv, observed 2026-05-20T10:58:13.533037Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.

source=pdf_text observed=2026-05-20T10:58:01.621488Z digest=sha256:94bf168e6609510faeae52c63b6eadee3ecd892e0281bd1ce3da144241fe027d

Observation 7caccb5e-28da-4de1-be3c-1ed7752b756c · inbound

Vision-OPD: Learning to See Fine Details for Multimodal LLMs via On-Policy Self-Distillation cites this paper.

Vision-OPD: Learning to See Fine Details for Multimodal LLMs via On-Policy Self-Distillation GLM-4.5V and GLM-4.1V-Thinking: Towards Versatile Multimodal Reasoning with Scalable Reinforcement Learning

Reference 16

Resolution
verified exact
local_arxiv, observed 2026-06-30T18:35:00.686256Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.

source=pdf_text observed=2026-06-30T18:28:21.605646Z digest=sha256:f0f139497b8ac9168287b6131e83b586151b7003dd5883dc668aca38a20180bf

Observation 2ace0ac1-81cd-473f-a744-b365884927b8 · inbound

Artifact-Bench: Evaluating MLLMs on Detecting and Assessing the Artifacts of AI-Generated Videos cites this paper.

Artifact-Bench: Evaluating MLLMs on Detecting and Assessing the Artifacts of AI-Generated Videos GLM-4.5V and GLM-4.1V-Thinking: Towards Versatile Multimodal Reasoning with Scalable Reinforcement Learning

Reference 12

Resolution
verified exact
local_arxiv, observed 2026-05-20T10:28:11.982339Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.

source=pdf_text observed=2026-05-20T10:26:30.661042Z digest=sha256:1ea86c78a4350f3bb1bc010eca911974ad54737d84640e933554fcdfbc92e41d

Observation 75780fc9-fb18-4b67-bc7d-502638b3bdc5 · inbound

HAVEN: Hierarchically Aligned Multimodal Benchmark for Unified Video Understanding cites this paper.

HAVEN: Hierarchically Aligned Multimodal Benchmark for Unified Video Understanding GLM-4.5V and GLM-4.1V-Thinking: Towards Versatile Multimodal Reasoning with Scalable Reinforcement Learning

Reference 19

Resolution
metadata mismatch
local_arxiv, observed 2026-05-20T07:43:08.497320Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.

source=pdf_text observed=2026-05-20T07:38:08.819186Z digest=sha256:964584d3ad89738ac6066293b82847d8d1bd52e93f7761f610aeeefe95334b05

Observation ddc24d59-a4eb-44f7-99de-0af6043e7b06 · inbound

Adaptive Probe-based Steering for Robust LLM Jailbreaking cites this paper.

Adaptive Probe-based Steering for Robust LLM Jailbreaking GLM-4.5V and GLM-4.1V-Thinking: Towards Versatile Multimodal Reasoning with Scalable Reinforcement Learning

Reference 41

Resolution
verified exact
local_arxiv, observed 2026-05-21T02:33:54.931423Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.

source=arxiv_source observed=2026-05-21T02:32:49.034790Z digest=sha256:ef60bb0df8c2e770f93de7daca485abdd215fe7680e5d7369b7959e24b27df5e

Observation 818e1f00-b120-4fed-9d59-746d3ea142f9 · inbound

ECUAS$_n$: A family of metrics for principled evaluation of uncertainty-augmented systems cites this paper.

ECUAS$_n$: A family of metrics for principled evaluation of uncertainty-augmented systems GLM-4.5V and GLM-4.1V-Thinking: Towards Versatile Multimodal Reasoning with Scalable Reinforcement Learning

Reference 79

Resolution
verified exact
local_arxiv, observed 2026-05-21T06:44:00.865303Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.

source=pdf_text observed=2026-05-21T06:42:18.133611Z digest=sha256:85be8ed53154554d893bd7dd6a2fdb38dbb0718a0a9ca8a82d272efaceaa07ce

Observation 853f7783-7202-4481-aa1e-3d86cb75e130 · inbound

ECUAS$_n$: A family of metrics for principled evaluation of uncertainty-augmented systems cites this paper.

ECUAS$_n$: A family of metrics for principled evaluation of uncertainty-augmented systems GLM-4.5V and GLM-4.1V-Thinking: Towards Versatile Multimodal Reasoning with Scalable Reinforcement Learning

Reference 79

Resolution
verified exact
local_arxiv, observed 2026-05-22T08:56:18.950186Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.

source=pdf_text observed=2026-05-22T08:55:07.587572Z digest=sha256:6567c7d0e635b4b836282136048817c8ec436241d03cca02ba2d4e6257671162

Observation 58ee22da-0a4e-4b43-9f9d-9050e36155ab · inbound

ECUAS$_n$: A family of metrics for principled evaluation of uncertainty-augmented systems cites this paper.

ECUAS$_n$: A family of metrics for principled evaluation of uncertainty-augmented systems GLM-4.5V and GLM-4.1V-Thinking: Towards Versatile Multimodal Reasoning with Scalable Reinforcement Learning

Reference 79

Resolution
verified exact
local_arxiv, observed 2026-06-30T17:54:57.697125Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.

source=pdf_text observed=2026-06-30T17:54:38.066497Z digest=sha256:274e0f2eb0c144da1c8bbfc8c43d343d0dbed271c37d93b04c8d4b818e72892b

Observation 66d90032-417f-4b92-b5d5-c14b681cca48 · inbound

Tracing the ongoing emergence of human-like reasoning in Large Language Models cites this paper.

Tracing the ongoing emergence of human-like reasoning in Large Language Models GLM-4.5V and GLM-4.1V-Thinking: Towards Versatile Multimodal Reasoning with Scalable Reinforcement Learning

Reference 72

Resolution
verified exact
local_arxiv, observed 2026-05-21T05:04:37.156476Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.

source=pdf_text observed=2026-05-21T05:04:30.829386Z digest=sha256:bd96a594b0aef79f23e21a0e7104972aa95b8d6514c86891da73fcb343d30601

Observation abfab5e9-1b23-4966-9193-c34f6f378029 · inbound

Frequency-Domain Regularized Adversarial Alignment for Transferable Attacks against Closed-Source MLLMs cites this paper.

Frequency-Domain Regularized Adversarial Alignment for Transferable Attacks against Closed-Source MLLMs GLM-4.5V and GLM-4.1V-Thinking: Towards Versatile Multimodal Reasoning with Scalable Reinforcement Learning

Reference 17

Resolution
verified exact
local_arxiv, observed 2026-05-22T01:25:52.627292Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.

source=pdf_text observed=2026-05-22T01:25:06.528845Z digest=sha256:64f655458f900b5428da75cb16c991e1c6220d03c2d2f44b7fd3a19393878fb9