Pith. sign in

Paper Citation Record · LEDGER

WebUIBench: A Comprehensive Benchmark for Evaluating Multimodal Large Language Models in WebUI-to-Code

As of 20 August 2026, this Paper Citation Record lists 35 of 35 outbound references and 2 inbound Pith citation observations for arXiv:2506.07818.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2506.07818 v1

Coverage vector

measured 35 of 35 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T05:27:50.207219Z

measured 37 of 37 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-20T06:33:59.587034+00:00

measured 2 of 2 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-15T23:21:13.398904Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-08-05T14:41:01.099787Z

Reference resolution

35 of 35 outbound references displayed

  • verified exact0
  • verified fuzzy1
  • unresolved34
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation fbf2e421-312d-4335-9823-f4a50e5a4cb2 · outbound

This paper cites Phi-3 Technical Report: A Highly Capable Language Model Locally on Your Phone.

WebUIBench: A Comprehensive Benchmark for Evaluating Multimodal Large Language Models in WebUI-to-Code Phi-3 Technical Report: A Highly Capable Language Model Locally on Your Phone

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-07T05:27:50.039856Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:27:50.039856Z digest=sha256:a62a146c9e295799d22e42584d8fec041444632ce76f073daa359bffe4cf6de6

Observation 671bea65-816c-4360-89b8-2fa35faaf777 · outbound

This paper cites The claude 3 model family: Opus, sonnet, haiku.

WebUIBench: A Comprehensive Benchmark for Evaluating Multimodal Large Language Models in WebUI-to-Code The claude 3 model family: Opus, sonnet, haiku

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-07T05:27:50.045581Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:27:50.045581Z digest=sha256:82b07f24f406f07d6691496f25d115fc8c815ab7d848ea244547ec93c7ce971e

Observation 7492d6bf-6c4b-476e-9406-bd7068abfb66 · outbound

This paper cites an unresolved cited work.

WebUIBench: A Comprehensive Benchmark for Evaluating Multimodal Large Language Models in WebUI-to-Code Unresolved cited work

Reference 4

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:27:50.902427Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-07T05:27:50.050471Z digest=sha256:ab87d7aaf2093412ebd6a77b6b8b64f6d1c3045a60d79fa9c4073d3a14622a76

Observation cf41afb1-0f8d-48df-9a40-536cffaa320b · outbound

This paper cites an unresolved cited work.

WebUIBench: A Comprehensive Benchmark for Evaluating Multimodal Large Language Models in WebUI-to-Code Unresolved cited work

Reference 5

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:27:50.889071Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-07T05:27:50.055521Z digest=sha256:5d43375d90cfbb0fa184607520bc8c3ed5da71d7b47f8bc32fd8cd7215ae304b

Observation 28a1da4d-8a62-400f-b4f3-171e136d64d8 · outbound

This paper cites A Survey on Evaluating Large Language Models in Code Generation Tasks.

WebUIBench: A Comprehensive Benchmark for Evaluating Multimodal Large Language Models in WebUI-to-Code A Survey on Evaluating Large Language Models in Code Generation Tasks

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-07T05:27:50.065829Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:27:50.065829Z digest=sha256:185df206d65b4bbda81e58e0cf95db93166dfb09a9a9ab3c9dc6609284d5af13

Observation 9f46b3ea-bc18-44f4-b119-946e3cd0bc83 · outbound

This paper cites Expanding Performance Boundaries of Open-Source Multimodal Models with Model, Data, and Test-Time Scaling.

WebUIBench: A Comprehensive Benchmark for Evaluating Multimodal Large Language Models in WebUI-to-Code Expanding Performance Boundaries of Open-Source Multimodal Models with Model, Data, and Test-Time Scaling

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-07T05:27:50.070819Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:27:50.070819Z digest=sha256:5d5021c2d913f20ea91edfe1096e40f85375b000ad28b5b32c4f784f8f5ae48b

Observation b921ebde-1312-446b-a08d-498f458ec84a · outbound

This paper cites an unresolved cited work.

WebUIBench: A Comprehensive Benchmark for Evaluating Multimodal Large Language Models in WebUI-to-Code Unresolved cited work

Reference 8

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:27:50.859503Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-07T05:27:50.075882Z digest=sha256:cc0506fa15d26dcc229115418411639107a2c6cea791b583a12e82e2e5abaf26

Observation cfd889bf-5ddb-45cc-9ddc-16eba5d4e6d9 · outbound

This paper cites an unresolved cited work.

WebUIBench: A Comprehensive Benchmark for Evaluating Multimodal Large Language Models in WebUI-to-Code Unresolved cited work

Reference 9

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:27:50.831773Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-07T05:27:50.085957Z digest=sha256:1fb42c1c0981f949ef4f287806f29173957c88d126709efa7a4fc991e1899bc0

Observation 67c5aa67-548f-46f2-8aab-7a1abb1231e9 · outbound

This paper cites ChatGLM: A Family of Large Language Models from GLM-130B to GLM-4 All Tools.

WebUIBench: A Comprehensive Benchmark for Evaluating Multimodal Large Language Models in WebUI-to-Code ChatGLM: A Family of Large Language Models from GLM-130B to GLM-4 All Tools

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-07T05:27:50.090673Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:27:50.090673Z digest=sha256:473dbb59430b1b57779321edc9e16939aaba0a7bac15dc7fd92126cdc038b3c1

Observation c09c10b2-0920-4810-bcaa-59cd4a2b1532 · outbound

This paper cites an unresolved cited work.

WebUIBench: A Comprehensive Benchmark for Evaluating Multimodal Large Language Models in WebUI-to-Code Unresolved cited work

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-07T05:27:50.096470Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:27:50.096470Z digest=sha256:748044c1ad7d40e3bef5b5cfda4c22369dbb7f5d00692a2c374077fc4adcf655

Observation 0412eb86-c49e-4a62-a859-f6b2f8b31154 · outbound

This paper cites GPT-4o System Card.

WebUIBench: A Comprehensive Benchmark for Evaluating Multimodal Large Language Models in WebUI-to-Code GPT-4o System Card

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-07T05:27:50.101164Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:27:50.101164Z digest=sha256:af6dc3f8162bcc9bc239d8ee39cfa486d8120caa81ed8fc1bd894dc3a35ce2a3

Observation edbb6bfc-40fb-4cad-b17c-263868b35414 · outbound

This paper cites an unresolved cited work.

WebUIBench: A Comprehensive Benchmark for Evaluating Multimodal Large Language Models in WebUI-to-Code Unresolved cited work

Reference 13

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:27:50.817435Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-07T05:27:50.106403Z digest=sha256:39de2c792ba298c7cc18f767e69d2224acbf6c690428511b6b0317830431de80

Observation 2ccfc924-d46b-43e4-90b0-abe1c759b149 · outbound

This paper cites an unresolved cited work.

WebUIBench: A Comprehensive Benchmark for Evaluating Multimodal Large Language Models in WebUI-to-Code Unresolved cited work

Reference 14

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:27:50.802724Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-07T05:27:50.116451Z digest=sha256:2f7d7a9c5729796efb0f5786bcb0d09bb5a3d6584188ce046c2430ccccb71a61

Observation 38bb9698-9827-4001-bcca-7e9a2deee6d4 · outbound

This paper cites Unlocking the conversion of Web Screenshots into HTML Code with the WebSight Dataset.

WebUIBench: A Comprehensive Benchmark for Evaluating Multimodal Large Language Models in WebUI-to-Code Unlocking the conversion of Web Screenshots into HTML Code with the WebSight Dataset

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-07T05:27:50.111566Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:27:50.111566Z digest=sha256:12d9e1ebf80b260acf3300d68107eaef56d06eda1e8e9e8a1536b6c5bf14b58a

Observation dd8e6060-44c9-4360-acc8-2401edbc7d62 · outbound

This paper cites A Survey on Multimodal Benchmarks: In the Era of Large AI Models.

WebUIBench: A Comprehensive Benchmark for Evaluating Multimodal Large Language Models in WebUI-to-Code A Survey on Multimodal Benchmarks: In the Era of Large AI Models

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-07T05:27:50.126766Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:27:50.126766Z digest=sha256:beec656d8149e465048c03f77ed8d86b38871efc28887eb837833b2dbb70b335

Observation 8eabf21d-8398-48fc-a512-5e5243d3ca70 · outbound

This paper cites SEED-Bench: Benchmarking Multimodal LLMs with Generative Comprehension.

WebUIBench: A Comprehensive Benchmark for Evaluating Multimodal Large Language Models in WebUI-to-Code SEED-Bench: Benchmarking Multimodal LLMs with Generative Comprehension

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-07T05:27:50.121929Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:27:50.121929Z digest=sha256:a7d913e0a87b86c05cb9262aab16133ede634deead421e34a48b2e8a4390f6c9

Observation 1f4d212c-0e72-4016-a0dd-0c62c367eb1e · outbound

This paper cites Ovis: Structural Embedding Alignment for Multimodal Large Language Model.

WebUIBench: A Comprehensive Benchmark for Evaluating Multimodal Large Language Models in WebUI-to-Code Ovis: Structural Embedding Alignment for Multimodal Large Language Model

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-07T05:27:50.136340Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:27:50.136340Z digest=sha256:eacd2917bdfaebe0b2237c1c9d8e1bacd95d3cf719d29c57fc3c3b7a80f6fdef

Observation 582d2760-34c3-4c13-a86b-8402941f4c61 · outbound

This paper cites an unresolved cited work.

WebUIBench: A Comprehensive Benchmark for Evaluating Multimodal Large Language Models in WebUI-to-Code Unresolved cited work

Reference 19

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:27:50.786294Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-07T05:27:50.131597Z digest=sha256:efb339095e3d9bf1db3e15632bfd738559c861ec42ec46883e1cb2a85ff95c7c

Observation 5de6a9c9-ac53-43b4-9682-0a7b347e2c0d · outbound

This paper cites an unresolved cited work.

WebUIBench: A Comprehensive Benchmark for Evaluating Multimodal Large Language Models in WebUI-to-Code Unresolved cited work

Reference 20

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:27:50.756555Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-07T05:27:50.145937Z digest=sha256:d444c2eba03c60ecfbc1690e94a9ee9e84958961b1bdf3e56ba97085a2a56c0c

Observation e4eea27e-3d9f-48db-863d-a2a5b6a3883f · outbound

This paper cites an unresolved cited work.

WebUIBench: A Comprehensive Benchmark for Evaluating Multimodal Large Language Models in WebUI-to-Code Unresolved cited work

Reference 21

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:27:50.771983Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-07T05:27:50.140917Z digest=sha256:ae0577dd2f9e56e568f887e205310bf41cf6ed7dcbf0040f3c71ebea06f8cbad

Observation d38f8e9b-6e87-41aa-a706-e357500e306c · outbound

This paper cites pix2code: Generating Code from a Graphical User Interface Screenshot.

WebUIBench: A Comprehensive Benchmark for Evaluating Multimodal Large Language Models in WebUI-to-Code pix2code: Generating Code from a Graphical User Interface Screenshot

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-07T05:27:50.155023Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:27:50.155023Z digest=sha256:2190c26e877491f54f7abe3e112a4fd519d0b402814a6007ae93ce4056106726

Observation 28525e44-676a-45cc-af5e-c69bddc1ec7a · outbound

This paper cites an unresolved cited work.

WebUIBench: A Comprehensive Benchmark for Evaluating Multimodal Large Language Models in WebUI-to-Code Unresolved cited work

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-07T05:27:50.150390Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:27:50.150390Z digest=sha256:14061dda703fea553ee41cbd69361863a6389f015109f58b0503350c8794d85f

Observation 9ab7e1df-22c4-438f-9bc2-65e36a1a914d · outbound

This paper cites an unresolved cited work.

WebUIBench: A Comprehensive Benchmark for Evaluating Multimodal Large Language Models in WebUI-to-Code Unresolved cited work

Reference 24

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:27:50.724914Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-07T05:27:50.164789Z digest=sha256:3e7bd93ed3fa2fe08fb1706d60ff01b40d8c6bdda187a04e46fa0e3cb38a61de

Observation 837eb6ba-c141-4f56-ba8c-a767dc3d279b · outbound

This paper cites an unresolved cited work.

WebUIBench: A Comprehensive Benchmark for Evaluating Multimodal Large Language Models in WebUI-to-Code Unresolved cited work

Reference 25

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:27:50.739755Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-07T05:27:50.159978Z digest=sha256:bfc3c2507f9b95735702c6c161182c59ca6a7a3a6c6b14230c7c13842a31376a

Observation 0052eaaa-fb5e-4114-977c-0ab8c2a20266 · outbound

This paper cites an unresolved cited work.

WebUIBench: A Comprehensive Benchmark for Evaluating Multimodal Large Language Models in WebUI-to-Code Unresolved cited work

Reference 26

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:27:50.710121Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-07T05:27:50.175237Z digest=sha256:7a0b656688d48c95ec7825e377eeb97250b4c9866c37a5969cdc969559e262d9

Observation df3b7bd5-453f-47de-beef-0f6e6d871a74 · outbound

This paper cites Design2Code: Benchmarking Multimodal Code Generation for Automated Front-End Engineering.

WebUIBench: A Comprehensive Benchmark for Evaluating Multimodal Large Language Models in WebUI-to-Code Design2Code: Benchmarking Multimodal Code Generation for Automated Front-End Engineering

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-07T05:27:50.170369Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:27:50.170369Z digest=sha256:6ac68391cbd878b538d48bd9a95868601143db4626b478376afca33bc3d7bb56

Observation 5f61bfa5-2bd9-43bd-8326-0db51f8349ee · outbound

This paper cites an unresolved cited work.

WebUIBench: A Comprehensive Benchmark for Evaluating Multimodal Large Language Models in WebUI-to-Code Unresolved cited work

Reference 28

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:27:50.695455Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-07T05:27:50.189437Z digest=sha256:d444157e65ed6c8bf3b4db34f0082c71191c2b38725f8a7dd20b4f3a1f6c558e

Observation 57a8284c-cc14-48a0-8c93-e485495524ae · outbound

This paper cites Gemini 1.5: Unlocking multimodal understanding across millions of tokens of context.

WebUIBench: A Comprehensive Benchmark for Evaluating Multimodal Large Language Models in WebUI-to-Code Gemini 1.5: Unlocking multimodal understanding across millions of tokens of context

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-07T05:27:50.180028Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:27:50.180028Z digest=sha256:9966f59147fed07aaab45e318bb54ffa152fac00967942708f61f681054a50c1

Observation 5f0e6284-b7ec-4559-85ce-9e1922115354 · outbound

This paper cites Qwen2-VL: Enhancing Vision-Language Model's Perception of the World at Any Resolution.

WebUIBench: A Comprehensive Benchmark for Evaluating Multimodal Large Language Models in WebUI-to-Code Qwen2-VL: Enhancing Vision-Language Model's Perception of the World at Any Resolution

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-07T05:27:50.184733Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:27:50.184733Z digest=sha256:67f96a5655f86dbc93067a9250364360d2b74f1365f7c1321158efb03f5ba759

Observation a77a577e-0620-4f7e-a949-71cbfd378f93 · outbound

This paper cites an unresolved cited work.

WebUIBench: A Comprehensive Benchmark for Evaluating Multimodal Large Language Models in WebUI-to-Code Unresolved cited work

Reference 31

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:27:50.680815Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-07T05:27:50.202973Z digest=sha256:198c6c9498cb8904710d88488184184e8286eb87120aa46148e1aadc6840434b

Observation 3a0af785-abc6-4c3d-9541-a5d56e14c59e · outbound

This paper cites MiniCPM-V: A GPT-4V Level MLLM on Your Phone.

WebUIBench: A Comprehensive Benchmark for Evaluating Multimodal Large Language Models in WebUI-to-Code MiniCPM-V: A GPT-4V Level MLLM on Your Phone

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-07T05:27:50.194019Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:27:50.194019Z digest=sha256:9f1e3ff78ceed724fe5b729a1e07c359a15ade7d000dec359731a1e5c8da83eb

Observation ad472f39-e63f-48cc-9896-f8f7cb521f2d · outbound

This paper cites Yi: Open Foundation Models by 01.AI.

WebUIBench: A Comprehensive Benchmark for Evaluating Multimodal Large Language Models in WebUI-to-Code Yi: Open Foundation Models by 01.AI

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-07T05:27:50.198375Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:27:50.198375Z digest=sha256:99a7e91cdd9ba867b421f2553512b64debbf05465d1949a1ddd4ed91c93dda6d

Observation 541a29be-d111-4864-844e-4c3dd74fd1d9 · outbound

This paper cites Web2Code: A Large-scale Webpage-to-Code Dataset and Evaluation Framework for Multimodal LLMs.

WebUIBench: A Comprehensive Benchmark for Evaluating Multimodal Large Language Models in WebUI-to-Code Web2Code: A Large-scale Webpage-to-Code Dataset and Evaluation Framework for Multimodal LLMs

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-07T05:27:50.207219Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:27:50.207219Z digest=sha256:0577692d4f86235e85ff02a0fd268ee7f2467bf9560f40d6e2621b2ab7a5906e

Observation 5e3e09cf-24ad-47ad-bca2-f61ddcdaa6f9 · outbound

This paper cites an unresolved cited work.

WebUIBench: A Comprehensive Benchmark for Evaluating Multimodal Large Language Models in WebUI-to-Code Unresolved cited work

Reference 2005

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:27:50.874432Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-07T05:27:50.060579Z digest=sha256:6da8508a3d452753d47ece62323f6857c0938fad92d2e8f5d08b5a33f90cacb5

Observation 09fb2401-0a1d-49fc-b72f-a2017d08e018 · outbound

This paper cites arXiv preprint.

WebUIBench: A Comprehensive Benchmark for Evaluating Multimodal Large Language Models in WebUI-to-Code arXiv preprint

Reference 2024

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:27:50.845803Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-07T05:27:50.080768Z digest=sha256:a8f3f06275bf821b49150eac2d060f163b9d2baea8e573f7d1d43303199dca9d

Pith citing papers

Observation ff465144-4dd3-4d8a-baff-d68ea3eb714b · inbound

Perception, Reason, Think, and Plan: A Survey on Large Multimodal Reasoning Models cites this paper.

Perception, Reason, Think, and Plan: A Survey on Large Multimodal Reasoning Models WebUIBench: A Comprehensive Benchmark for Evaluating Multimodal Large Language Models in WebUI-to-Code

Reference 284

Resolution
unresolved
no resolver link, observed 2026-08-15T23:21:13.398904Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T23:21:13.398904Z digest=sha256:5657d2a20ebfc04704d34a0b65bcc34a32289f6523b017e5e13b45955b882e4f

Observation 9aede403-8318-4a05-a146-1b6eace7ca0d · inbound

Pattern over Pixels: Measuring Pattern Completion Bias in Multimodal Code Generation cites this paper.

Pattern over Pixels: Measuring Pattern Completion Bias in Multimodal Code Generation WebUIBench: A Comprehensive Benchmark for Evaluating Multimodal Large Language Models in WebUI-to-Code

Reference 23

Resolution
verified exact
local_arxiv, observed 2026-08-05T14:41:01.107702Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-05T14:41:00.211355Z digest=sha256:54cc9c1b6b65e2a08688fff6439024ff2d8341452f4ed91df1b5335857df3273