Pith. sign in

Paper Citation Record · LEDGER

GLM-5V-Turbo: Toward a Native Foundation Model for Multimodal Agents

As of 5 August 2026, this Paper Citation Record lists 50 of 50 outbound references and 11 inbound Pith citation observations for arXiv:2604.26752.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2604.26752 v3

Coverage vector

measured 50 of 50 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-05-13T07:28:45.811192Z

measured 61 of 61 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-05T06:32:48.257954+00:00

measured 11 of 11 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-05T18:29:09.689924Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-07-04T10:09:45.391749Z

Reference resolution

50 of 50 outbound references displayed

  • verified exact15
  • verified fuzzy16
  • unresolved18
  • parse uncertain1
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 7b9a8539-f951-4d70-9963-5d5d3a9bde4e · outbound

This paper cites an unresolved cited work.

GLM-5V-Turbo: Toward a Native Foundation Model for Multimodal Agents Unresolved cited work

Reference 1

Resolution
unresolved
raw_fallback, observed 2026-05-13T08:22:31.829238Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-13T07:28:45.811192Z digest=sha256:602637368fb2aa9967e93a0120193481fa9e96a928ca3f6dd9dae472ce3512ee

Observation c1e945f5-2814-42fd-be52-822e0eb00f7c · outbound

This paper cites an unresolved cited work.

GLM-5V-Turbo: Toward a Native Foundation Model for Multimodal Agents Unresolved cited work

Reference 2

Resolution
unresolved
raw_fallback, observed 2026-05-13T08:22:31.833444Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-13T07:28:45.811192Z digest=sha256:df5121b6df346df3a69ea35c1bfd540a07f1b50f2b467b4d5362daf272f12f1d

Observation 65b6a156-f19a-49b0-9a13-9780d42058cb · outbound

This paper cites Claude code: Ai-powered coding assistant.

GLM-5V-Turbo: Toward a Native Foundation Model for Multimodal Agents Claude code: Ai-powered coding assistant

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-05-13T08:22:31.817731Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-13T07:28:45.811192Z digest=sha256:d3f78748d2164e7b68650f8d2d9e7b78ca160d1e4ec89b5ff82bc7dac944921c

Observation 7a033d59-b5ab-427b-9e86-075d527e4a53 · outbound

This paper cites Introducing claude opus 4.6.

GLM-5V-Turbo: Toward a Native Foundation Model for Multimodal Agents Introducing claude opus 4.6

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-05-13T08:22:31.822621Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-13T07:28:45.811192Z digest=sha256:9ddcd896e088df2a500161b99e8a03ceca73f6002f9134245d425e85da3157a3

Observation 548ef287-21c9-48d0-bed3-a8582483d2e9 · outbound

This paper cites Seed2.0 model card: Towards intelligence frontier for real-world complexity.

GLM-5V-Turbo: Toward a Native Foundation Model for Multimodal Agents Seed2.0 model card: Towards intelligence frontier for real-world complexity

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-05-13T08:22:31.929610Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-13T07:28:45.811192Z digest=sha256:679d57fa00cd4f1c77df56765d9771c17999ae157ff09048f09feb7c07d66742

Observation ccd375ba-116a-4bfe-b769-a790e05b02b3 · outbound

This paper cites PointArena: Probing Multimodal Grounding Through Language-Guided Pointing.

GLM-5V-Turbo: Toward a Native Foundation Model for Multimodal Agents PointArena: Probing Multimodal Grounding Through Language-Guided Pointing

Reference 6

Resolution
verified exact
arxiv_id, observed 2026-05-13T07:32:30.501953Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-13T07:28:45.811192Z digest=sha256:a872ae89fa8be035f7462c154f8e561d93c7bc1a5fef9188e59b1e18f8461eac

Observation f14c9e75-5b69-4420-8b4a-ceb15b4045f8 · outbound

This paper cites Cheng, W.

GLM-5V-Turbo: Toward a Native Foundation Model for Multimodal Agents Cheng, W

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-05-13T08:22:31.976624Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-13T07:28:45.811192Z digest=sha256:6ea6250a94a2fae53ff0d1cfade79e9a22c0fb430687c2924012cb8b0101c157

Observation 39e2e7ab-185a-4c36-a54e-3de3e5657c33 · outbound

This paper cites an unresolved cited work.

GLM-5V-Turbo: Toward a Native Foundation Model for Multimodal Agents Unresolved cited work

Reference 8

Resolution
unresolved
raw_fallback, observed 2026-05-13T08:22:31.907499Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-13T07:28:45.811192Z digest=sha256:d74d55f9fac43ccd0cb30da00ec8525769582a0d71110d5806e9df914cb7d9f9

Observation 583f0bc7-6720-470c-bd68-ca9fa9f61f47 · outbound

This paper cites Advancing vision-language models in front-end development via data synthesis.

GLM-5V-Turbo: Toward a Native Foundation Model for Multimodal Agents Advancing vision-language models in front-end development via data synthesis

Reference 9

Resolution
verified exact
arxiv_id, observed 2026-05-13T07:32:30.478163Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-13T07:28:45.811192Z digest=sha256:02901f55e992e7cc135051f8dcac11bbfd8448eab3c356401295bfcf6d3eeda4

Observation 85503728-ddb0-469f-9871-fbf347707b04 · outbound

This paper cites an unresolved cited work.

GLM-5V-Turbo: Toward a Native Foundation Model for Multimodal Agents Unresolved cited work

Reference 10

Resolution
unresolved
raw_fallback, observed 2026-05-13T08:22:31.961224Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-13T07:28:45.811192Z digest=sha256:a97302b839d2a92d7af52c1d6e96a810cda1a98ab6b47fd99da07ba972f6917a

Observation 2cf6ea7a-5137-4d1b-b155-fb6adbbf446c · outbound

This paper cites Better & Faster Large Language Models via Multi-token Prediction.

GLM-5V-Turbo: Toward a Native Foundation Model for Multimodal Agents Better & Faster Large Language Models via Multi-token Prediction

Reference 11

Resolution
verified exact
arxiv_id, observed 2026-05-16T12:26:09.884823Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-13T07:28:45.811192Z digest=sha256:3b67926f0df696553c3863d7f0ffecaa03172c5ed514f6085367592479bd7fe5

Observation 788a719c-2428-455b-b46a-03f1e11b1812 · outbound

This paper cites The latest updates for Deep Research in Gemini.

GLM-5V-Turbo: Toward a Native Foundation Model for Multimodal Agents The latest updates for Deep Research in Gemini

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-05-13T08:22:31.925182Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-13T07:28:45.811192Z digest=sha256:f8f5f580babe8352cd4920dccb70051329e77fdf3818e2f1d697a9945149a2f6

Observation 7d95dfef-c728-4961-99a9-31136c0df393 · outbound

This paper cites WebVoyager: Building an End-to-End Web Agent with Large Multimodal Models.

GLM-5V-Turbo: Toward a Native Foundation Model for Multimodal Agents WebVoyager: Building an End-to-End Web Agent with Large Multimodal Models

Reference 13

Resolution
verified exact
arxiv_id, observed 2026-05-15T22:43:35.479641Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-13T07:28:45.811192Z digest=sha256:8d6319ee3f30bbc52be02439b418bdeb4cc17d804d6a8b846c8789f4421e10aa

Observation 8044785c-a366-47f7-8a06-b3906713e3a0 · outbound

This paper cites Vision2Web: A Hierarchical Benchmark for Visual Website Development with Agent Verification.

GLM-5V-Turbo: Toward a Native Foundation Model for Multimodal Agents Vision2Web: A Hierarchical Benchmark for Visual Website Development with Agent Verification

Reference 14

Resolution
verified exact
arxiv_id, observed 2026-07-21T02:21:34.388156Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-13T07:28:45.811192Z digest=sha256:ccb4570c53e803ae006b0e30643b59a9e0b55dbf5ade06018b01dbab4bbb130c

Observation cccd50e1-4b82-4fca-8413-fdd5073d306b · outbound

This paper cites Henry, P.

GLM-5V-Turbo: Toward a Native Foundation Model for Multimodal Agents Henry, P

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-05-13T08:22:31.921453Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-13T07:28:45.811192Z digest=sha256:107cc119448dee83dc5a6b227ef60167c6abc686a32a058e8c1d1fdfadb74074

Observation d3bab311-fcc7-45ac-b287-6c2d332cdf6c · outbound

This paper cites an unresolved cited work.

GLM-5V-Turbo: Toward a Native Foundation Model for Multimodal Agents Unresolved cited work

Reference 16

Resolution
unresolved
raw_fallback, observed 2026-05-13T08:22:31.899443Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-13T07:28:45.811192Z digest=sha256:687110b3ee65c4d8ff9f52d83bbef9cf4e0d708d8bd0b3bc32089dc5fde7d260

Observation 1c35be63-ea6c-4b22-b0e9-257358ae17b7 · outbound

This paper cites Jacovi, A.

GLM-5V-Turbo: Toward a Native Foundation Model for Multimodal Agents Jacovi, A

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-05-13T08:22:31.937846Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-13T07:28:45.811192Z digest=sha256:fcaa2c4f4889fe7588e5513cc1d386f4856bd488edcdfbc541e98bb619cb0644

Observation c4be0e21-1769-403b-b0ef-76bcf45e8a62 · outbound

This paper cites an unresolved cited work.

GLM-5V-Turbo: Toward a Native Foundation Model for Multimodal Agents Unresolved cited work

Reference 18

Resolution
unresolved
raw_fallback, observed 2026-05-13T08:22:31.933589Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-13T07:28:45.811192Z digest=sha256:eac3dc50ce77b855738cfd77edfbe9de71d6da22cbe5d9b39c8f81be78cc14b0

Observation 04ced225-e5a5-4043-81f2-d4b0bc912c05 · outbound

This paper cites MMSearch: Benchmarking the Potential of Large Models as Multi-modal Search Engines.

GLM-5V-Turbo: Toward a Native Foundation Model for Multimodal Agents MMSearch: Benchmarking the Potential of Large Models as Multi-modal Search Engines

Reference 20

Resolution
verified exact
arxiv_id, observed 2026-05-13T07:32:30.516597Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-13T07:28:45.811192Z digest=sha256:049df8a6f3220bc089c2e79e9686e22a5a219e54a8271f93ac0d51660ee88b29

Observation 72e1e3df-a385-4b94-8552-71b7a38dc9b4 · outbound

This paper cites SWE-bench: Can Language Models Resolve Real-World GitHub Issues?.

GLM-5V-Turbo: Toward a Native Foundation Model for Multimodal Agents SWE-bench: Can Language Models Resolve Real-World GitHub Issues?

Reference 21

Resolution
verified exact
local_arxiv, observed 2026-05-13T07:32:30.511303Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-13T07:28:45.811192Z digest=sha256:8669e252f3aa38e6c9adc48739466b0bbc0d1ded018eefddbcd89a560a9f6d4e

Observation d1e75dff-b34f-4adb-88c4-6707198e5b70 · outbound

This paper cites Jordan et al.

GLM-5V-Turbo: Toward a Native Foundation Model for Multimodal Agents Jordan et al

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-05-13T08:22:31.873355Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-13T07:28:45.811192Z digest=sha256:4d0c9542f3c2827e355a65f7e286400bc63c8bda0c5c71b0f6bdd70b513854ba

Observation 63327540-69cf-474d-85ca-5ae7722d5909 · outbound

This paper cites Karpathy.

GLM-5V-Turbo: Toward a Native Foundation Model for Multimodal Agents Karpathy

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-05-13T08:22:31.887491Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-13T07:28:45.811192Z digest=sha256:ccbcd220d6709a04f09293460d254716ebc673d1da2cd40a5cd6844a1d54dd3c

Observation 8032fc12-1d0f-42fc-b297-06912179d0d2 · outbound

This paper cites Kazemzadeh, V.

GLM-5V-Turbo: Toward a Native Foundation Model for Multimodal Agents Kazemzadeh, V

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-05-13T08:22:31.948392Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-13T07:28:45.811192Z digest=sha256:b9593d321c59b16f4e7491cf57f121028a8cc7b592793a9be78707eb116dc109

Observation 74a82e5c-e2e3-49fc-aed2-5bae51e21725 · outbound

This paper cites an unresolved cited work.

GLM-5V-Turbo: Toward a Native Foundation Model for Multimodal Agents Unresolved cited work

Reference 25

Resolution
unresolved
raw_fallback, observed 2026-05-13T08:22:31.891311Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-13T07:28:45.811192Z digest=sha256:02959b298a0ad6f15b5a972410657e797e3c79fdb12a0f4d95590878f296dc41

Observation 97d527d8-2777-453e-aa2c-c2a68924178e · outbound

This paper cites an unresolved cited work.

GLM-5V-Turbo: Toward a Native Foundation Model for Multimodal Agents Unresolved cited work

Reference 26

Resolution
unresolved
raw_fallback, observed 2026-05-13T08:22:31.915735Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-13T07:28:45.811192Z digest=sha256:7c01cdb61b43c5754049480d682eff4a2ad7a8644e2c41ebbc6699622d4816dc

Observation d940ec69-a774-429d-85b8-647a5ce27db8 · outbound

This paper cites MathVista: Evaluating Mathematical Reasoning of Foundation Models in Visual Contexts.

GLM-5V-Turbo: Toward a Native Foundation Model for Multimodal Agents MathVista: Evaluating Mathematical Reasoning of Foundation Models in Visual Contexts

Reference 27

Resolution
verified exact
local_arxiv, observed 2026-05-13T07:32:30.492606Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-13T07:28:45.811192Z digest=sha256:199513b6cd2339e1dd1252b2c3deeb90fd90c4b8bba3e6f2883bd2392d3535f8

Observation 020f7443-44c9-4e68-8d45-cf64e653bc68 · outbound

This paper cites Introducing deep research.

GLM-5V-Turbo: Toward a Native Foundation Model for Multimodal Agents Introducing deep research

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-05-13T08:22:31.878966Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-13T07:28:45.811192Z digest=sha256:06152bf6f76ef2aab11ef74ab5c8c13da505ff598aa8558c2cd5edcf2216dc94

Observation efe4ac10-86f7-4d1a-860b-858c42d38e66 · outbound

This paper cites Introducing gpt-5.4.

GLM-5V-Turbo: Toward a Native Foundation Model for Multimodal Agents Introducing gpt-5.4

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-05-13T08:22:31.944672Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-13T07:28:45.811192Z digest=sha256:41e4d511d534e2e6a450c47298f790ab6b0cc18bbacd0162afb54d78a5d9a47c

Observation a75a3a08-f7ce-4566-bbe8-2f24627a0627 · outbound

This paper cites an unresolved cited work.

GLM-5V-Turbo: Toward a Native Foundation Model for Multimodal Agents Unresolved cited work

Reference 30

Resolution
parse uncertain
raw_fallback, observed 2026-05-13T08:22:31.856442Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-13T07:28:45.811192Z digest=sha256:273188fd0428eadbfd8fc7eaeb8723c8896e0283de12c4ccabccb8c4b4685ac6

Observation 39674ea6-dbf5-4041-9617-34620976d8b1 · outbound

This paper cites Openclaw.

GLM-5V-Turbo: Toward a Native Foundation Model for Multimodal Agents Openclaw

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-05-13T08:22:31.849831Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-13T07:28:45.811192Z digest=sha256:0bf42381c722a8b9a1c7abfca9d2ad6f7ad055ee3d62977391cd0eff32d8e439

Observation f68eaf73-8d46-4d06-b138-ace89458a825 · outbound

This paper cites AndroidWorld: A Dynamic Benchmarking Environment for Autonomous Agents.

GLM-5V-Turbo: Toward a Native Foundation Model for Multimodal Agents AndroidWorld: A Dynamic Benchmarking Environment for Autonomous Agents

Reference 32

Resolution
verified exact
arxiv_id, observed 2026-05-13T12:06:13.928391Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-13T07:28:45.811192Z digest=sha256:8750ec3c30ad29188255cf27264ae00fdbfd7e946580b1061c5cda6183e7255a

Observation 56360865-4f98-418e-a710-d62142d06158 · outbound

This paper cites an unresolved cited work.

GLM-5V-Turbo: Toward a Native Foundation Model for Multimodal Agents Unresolved cited work

Reference 33

Resolution
unresolved
raw_fallback, observed 2026-05-13T08:22:31.883517Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-13T07:28:45.811192Z digest=sha256:ed95a75e1d370b01470ae812da0e1e85003bc15f7f823755124d59420f0d99b4

Observation 001477c9-59e8-4d72-97d2-37efd0f58d59 · outbound

This paper cites DINOv3.

GLM-5V-Turbo: Toward a Native Foundation Model for Multimodal Agents DINOv3

Reference 34

Resolution
verified exact
local_arxiv, observed 2026-05-13T07:32:30.521833Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-13T07:28:45.811192Z digest=sha256:1d613b0edc787ef30cc3b36215a2610e0c9d619f9a13c7b3bd0233fb895354b3

Observation 00dad2d1-1596-4bc3-b8d7-f5efa9e4c75c · outbound

This paper cites an unresolved cited work.

GLM-5V-Turbo: Toward a Native Foundation Model for Multimodal Agents Unresolved cited work

Reference 35

Resolution
unresolved
raw_fallback, observed 2026-05-13T08:22:31.912002Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-13T07:28:45.811192Z digest=sha256:9bb4f4380534f540e373c2a798f0a1061579e38bb1cd2bb0003831a277370613

Observation 69b90f8b-7e84-472a-a570-fcc1a3d0c9b4 · outbound

This paper cites Steinberger.

GLM-5V-Turbo: Toward a Native Foundation Model for Multimodal Agents Steinberger

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-05-13T08:22:31.956681Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-13T07:28:45.811192Z digest=sha256:0747773138d1a9669122452440f5343c8a46750598851ebb9c4d701f37f90c56

Observation 3cf1a700-f4ee-48a7-9881-33b27a616b5a · outbound

This paper cites arXiv preprint arXiv:2508.21475 , year=.

GLM-5V-Turbo: Toward a Native Foundation Model for Multimodal Agents arXiv preprint arXiv:2508.21475 , year=

Reference 37

Resolution
verified exact
arxiv_id, observed 2026-05-13T07:32:30.497430Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-13T07:28:45.811192Z digest=sha256:939939c17fae607bcacf1112a61094bd812e7b4c14a407db6471ba5b6236ba3d

Observation 9b89c985-455c-4915-9bc6-edad5de5e81f · outbound

This paper cites an unresolved cited work.

GLM-5V-Turbo: Toward a Native Foundation Model for Multimodal Agents Unresolved cited work

Reference 38

Resolution
unresolved
raw_fallback, observed 2026-05-13T08:22:31.941305Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-13T07:28:45.811192Z digest=sha256:91b48b6f2743d1b19541fad6670ab702e89479e895d696a03432408639ac074f

Observation 56d8f400-7760-4ecb-b2e2-6b6a92b3de0f · outbound

This paper cites an unresolved cited work.

GLM-5V-Turbo: Toward a Native Foundation Model for Multimodal Agents Unresolved cited work

Reference 39

Resolution
unresolved
raw_fallback, observed 2026-05-13T08:22:31.866681Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-13T07:28:45.811192Z digest=sha256:ff6d65ee1ad1ea942f3cfbd85ffd7db32262e91ce7412eb01b804a07a584b10a

Observation 2fcd0186-b73f-4f1e-8cae-a3549625a314 · outbound

This paper cites an unresolved cited work.

GLM-5V-Turbo: Toward a Native Foundation Model for Multimodal Agents Unresolved cited work

Reference 40

Resolution
unresolved
raw_fallback, observed 2026-05-13T08:22:31.965187Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-13T07:28:45.811192Z digest=sha256:7e315790444eccd5b3192ec6c2d65fac13f5be8d08b744066ad90f92fd9a08e9

Observation 7c2f6dea-184e-48f9-9cbb-69ddb766d874 · outbound

This paper cites SigLIP 2: Multilingual Vision-Language Encoders with Improved Semantic Understanding, Localization, and Dense Features.

GLM-5V-Turbo: Toward a Native Foundation Model for Multimodal Agents SigLIP 2: Multilingual Vision-Language Encoders with Improved Semantic Understanding, Localization, and Dense Features

Reference 41

Resolution
verified exact
local_arxiv, observed 2026-05-13T07:32:30.482826Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-13T07:28:45.811192Z digest=sha256:b65dd3262949452c9ff4452ca03107d3b7541ced681caa93cd1a60b22dd74a55

Observation 7b904547-73b8-42aa-9c62-72af4e4b2476 · outbound

This paper cites an unresolved cited work.

GLM-5V-Turbo: Toward a Native Foundation Model for Multimodal Agents Unresolved cited work

Reference 42

Resolution
unresolved
raw_fallback, observed 2026-05-13T08:22:31.895168Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-13T07:28:45.811192Z digest=sha256:defdd4fa18e09fdb82b5e6cbbf5c44639b067028b939a32c1ec23c364811e757

Observation 05ff0976-06b0-42d7-9706-6cf44db719bf · outbound

This paper cites Wu and S.

GLM-5V-Turbo: Toward a Native Foundation Model for Multimodal Agents Wu and S

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-05-13T08:22:31.951611Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-13T07:28:45.811192Z digest=sha256:ef23646685e1275e567740b74a585776670ec620b43236e303beea2939c16d2c

Observation c8219ebd-e58b-426c-b3f8-886806d4afdd · outbound

This paper cites LogicVista: Multimodal LLM Logical Reasoning Benchmark in Visual Contexts.

GLM-5V-Turbo: Toward a Native Foundation Model for Multimodal Agents LogicVista: Multimodal LLM Logical Reasoning Benchmark in Visual Contexts

Reference 44

Resolution
verified exact
arxiv_id, observed 2026-05-16T02:47:26.275168Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-13T07:28:45.811192Z digest=sha256:1924ac047c35e7dd194903f5b4a47ec82a9a61a4dbee4ad43a3e1dec6e9caa27

Observation a5e7b4c9-6cb9-40b5-a2c9-03c3a879c87e · outbound

This paper cites an unresolved cited work.

GLM-5V-Turbo: Toward a Native Foundation Model for Multimodal Agents Unresolved cited work

Reference 45

Resolution
unresolved
raw_fallback, observed 2026-05-13T08:22:31.972715Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-13T07:28:45.811192Z digest=sha256:2dae204bd5a88efdaaa5d7e3f22cbba93cb81d6ce62d621868128193789bb1f3

Observation 23c4db89-6517-41ff-8a43-10afd7ea51fc · outbound

This paper cites an unresolved cited work.

GLM-5V-Turbo: Toward a Native Foundation Model for Multimodal Agents Unresolved cited work

Reference 46

Resolution
unresolved
raw_fallback, observed 2026-05-13T08:22:31.841580Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-13T07:28:45.811192Z digest=sha256:4bef197c5fcb48c212113c6c162c7d783a4cdc381674e8a112bb444f58d77214

Observation 08777abb-c136-4bc5-be92-5ce72b479e1f · outbound

This paper cites Claw-Eval: Towards Trustworthy Evaluation of Autonomous Agents.

GLM-5V-Turbo: Toward a Native Foundation Model for Multimodal Agents Claw-Eval: Towards Trustworthy Evaluation of Autonomous Agents

Reference 47

Resolution
verified exact
local_arxiv, observed 2026-05-13T07:32:30.468067Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-13T07:28:45.811192Z digest=sha256:68f7a3b605aa187fa1b77c2929535a3afdb4681a22cf2039b6a2273ca0ca0561

Observation be2727da-8678-4532-9815-1cec332654e9 · outbound

This paper cites an unresolved cited work.

GLM-5V-Turbo: Toward a Native Foundation Model for Multimodal Agents Unresolved cited work

Reference 48

Resolution
unresolved
raw_fallback, observed 2026-05-13T08:22:31.862949Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-13T07:28:45.811192Z digest=sha256:4f806b37552cea4e1a7eae36d826c131e75d5425b1ba37fb129ba0a4b9b044cc

Observation 383e5d36-29e3-4326-bd79-a3869a31fec2 · outbound

This paper cites an unresolved cited work.

GLM-5V-Turbo: Toward a Native Foundation Model for Multimodal Agents Unresolved cited work

Reference 49

Resolution
unresolved
raw_fallback, observed 2026-05-13T08:22:31.969030Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-13T07:28:45.811192Z digest=sha256:ba0851d68824883eb715babd387fe073731986691d9a7bab966b4db5da08fcff

Observation 70f5e44f-5313-4d0e-87d2-6cf20fcdcb31 · outbound

This paper cites GLM-5: from Vibe Coding to Agentic Engineering.

GLM-5V-Turbo: Toward a Native Foundation Model for Multimodal Agents GLM-5: from Vibe Coding to Agentic Engineering

Reference 50

Resolution
verified exact
local_arxiv, observed 2026-05-13T07:32:30.526949Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-13T07:28:45.811192Z digest=sha256:946d75e9db81f1fc51d359d76ac8260a69d00720b56977bba20de71333589e81

Observation 4ef9b497-3b47-465f-ad10-88be1b092388 · outbound

This paper cites About Brand.

GLM-5V-Turbo: Toward a Native Foundation Model for Multimodal Agents About Brand

Reference 51

Resolution
verified fuzzy
raw_fallback, observed 2026-05-13T08:22:31.845986Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-13T07:28:45.811192Z digest=sha256:1628f1332926bb0fb7bf2237cd8a006ead6b91554038db2456ee2d7af00d3c17

Pith citing papers

Observation 30a7c1fb-bb69-4545-af51-6199506411ee · inbound

FIKA-Bench: From Fine-grained Recognition to Fine-Grained Knowledge Acquisition cites this paper.

FIKA-Bench: From Fine-grained Recognition to Fine-Grained Knowledge Acquisition GLM-5V-Turbo: Toward a Native Foundation Model for Multimodal Agents

Reference 14

Resolution
verified exact
local_arxiv, observed 2026-05-14T20:42:58.694201Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-14T20:25:36.746335Z digest=sha256:ad7b58934295e30a7bc22c845f458ea717184f9e16d0d355e5a323d5c7d98273

Observation e1d0bad3-e998-42c7-921b-d2ad3e9686bc · inbound

FIKA-Bench: From Fine-grained Recognition to Fine-Grained Knowledge Acquisition cites this paper.

FIKA-Bench: From Fine-grained Recognition to Fine-Grained Knowledge Acquisition GLM-5V-Turbo: Toward a Native Foundation Model for Multimodal Agents

Reference 14

Resolution
verified exact
local_arxiv, observed 2026-05-20T22:03:47.187415Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-20T22:02:19.717335Z digest=sha256:505b42cc7522d6d144ec440ca6cc9d403045f3ca6747e794d81bbe23a0cec65d

Observation 724b1f36-3924-465b-8c8d-8c3e27935671 · inbound

VideoSeeker: Incentivizing Instance-level Video Understanding via Native Agentic Tool Invocation cites this paper.

VideoSeeker: Incentivizing Instance-level Video Understanding via Native Agentic Tool Invocation GLM-5V-Turbo: Toward a Native Foundation Model for Multimodal Agents

Reference 15

Resolution
verified exact
local_arxiv, observed 2026-05-20T19:38:56.204131Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-20T19:37:09.244578Z digest=sha256:62bc2f8d067e4d61111a01a050e209860e13efc8a4ee88b85ed004364ef1ac62

Observation 58e58408-69c2-4b73-bdf8-af4a224bb64d · inbound

SpatialAct: Probing Spatial Reasoning-to-Action Capabilities of VLM Agents in 3D Scenes cites this paper.

SpatialAct: Probing Spatial Reasoning-to-Action Capabilities of VLM Agents in 3D Scenes GLM-5V-Turbo: Toward a Native Foundation Model for Multimodal Agents

Reference 6

Resolution
verified exact
local_arxiv, observed 2026-06-28T22:52:45.651730Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-06-28T22:47:46.542267Z digest=sha256:ea621f7e02cd14ed42bf0c479f5744a5765c24e558aaef904d857dd65140566e

Observation 94c8a7d4-22c8-44fb-a145-7a4b507290c7 · inbound

HLL: Can Agents Cross Humanity's Last Line of Verification? cites this paper.

HLL: Can Agents Cross Humanity's Last Line of Verification? GLM-5V-Turbo: Toward a Native Foundation Model for Multimodal Agents

Reference 22

Resolution
verified exact
local_arxiv, observed 2026-07-01T22:56:19.632650Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-06-28T14:57:57.218669Z digest=sha256:76d1d88ec502211aef411354b4d364abc09295ac5acd198a2e20d9a30e345f72

Observation fe227577-00b1-4251-9126-b9ecf423a76a · inbound

Imagine Before You Predict: Interleaved Latent Visual Reasoning for Video Event Prediction cites this paper.

Imagine Before You Predict: Interleaved Latent Visual Reasoning for Video Event Prediction GLM-5V-Turbo: Toward a Native Foundation Model for Multimodal Agents

Reference 50

Resolution
metadata mismatch
local_arxiv, observed 2026-07-02T11:56:55.439810Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-06-28T02:46:50.373450Z digest=sha256:5d1f4050417f8fc6959062db7d05be11bc0af6fe76fb1177223788e97f94b93e

Observation 32b6dc1d-40a6-4d08-98ea-c2237b1e8d06 · inbound

Each Judge Its Own Yardstick: Discovering Per-VLM Taxonomies for Physical Video Evaluation cites this paper.

Each Judge Its Own Yardstick: Discovering Per-VLM Taxonomies for Physical Video Evaluation GLM-5V-Turbo: Toward a Native Foundation Model for Multimodal Agents

Reference 47

Resolution
metadata mismatch
local_arxiv, observed 2026-07-04T10:09:45.392896Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-06-26T09:04:23.965554Z digest=sha256:853db94cfa393812fc2719ea00c94c792eb25f4e86bf970f8e7e7f5a2531e147

Observation 50ddd35e-7d1a-4186-833c-81f995e28d4a · inbound

IoU-PD: IoU-Aware Privileged Distillation for Visual Grounding with Multimodal Large Language Models cites this paper.

IoU-PD: IoU-Aware Privileged Distillation for Visual Grounding with Multimodal Large Language Models GLM-5V-Turbo: Toward a Native Foundation Model for Multimodal Agents

Reference 118

Resolution
unresolved
no resolver link, observed 2026-08-01T22:32:28.819129Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T22:32:28.819129Z digest=sha256:67a335229ecf0cd31cec187f8f10b60981e554ade092ad821168c786d1d6b6ff

Observation ee087e6f-e510-400a-89b9-aab1a6c49c5c · inbound

IoU-PD: IoU-Aware Privileged Distillation for Visual Grounding with Multimodal Large Language Models cites this paper.

IoU-PD: IoU-Aware Privileged Distillation for Visual Grounding with Multimodal Large Language Models GLM-5V-Turbo: Toward a Native Foundation Model for Multimodal Agents

Reference 118

Resolution
unresolved
no resolver link, observed 2026-08-04T04:21:09.624375Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T04:21:09.624375Z digest=sha256:9a120c148eea121bd5e98e58d8191c396d9f22d1f44049c9b028467b4db8d530

Observation 1b7798bc-7e88-48d5-b255-6f7969674115 · inbound

DataClawEval: A Benchmark for Data Engineering Agents in Real Industrial Harness cites this paper.

DataClawEval: A Benchmark for Data Engineering Agents in Real Industrial Harness GLM-5V-Turbo: Toward a Native Foundation Model for Multimodal Agents

Reference 10

Resolution
unresolved
no resolver link, observed 2026-07-31T19:54:08.429432Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T19:54:08.429432Z digest=sha256:fb9044126a360413c281e79d28d27dcbaa011a0bc05129ff25edc69c8dbc2927

Observation 646ae864-2019-4381-ab51-528c0f5a518f · inbound

MT-Web2Code: Benchmarking Coding Agents on Multi-Turn Regional Reconstruction and Localized Modification cites this paper.

MT-Web2Code: Benchmarking Coding Agents on Multi-Turn Regional Reconstruction and Localized Modification GLM-5V-Turbo: Toward a Native Foundation Model for Multimodal Agents

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-05T18:29:09.689924Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T18:29:09.689924Z digest=sha256:4264fdcdfd6934a5b87e5f8d438a2989a380d96a3eae3b2c9c1603d27b6aca82