Pith. sign in

Paper Citation Record · LEDGER

WebUIBench: A Comprehensive Benchmark for Evaluating Multimodal Large Language Models in WebUI-to-Code

As of 8 August 2026, this Paper Citation Record lists 35 of 35 outbound references and 1 inbound Pith citation observation for arXiv:2506.07818.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2506.07818 v1

Coverage vector

measured 35 of 35 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T05:27:50.207219Z

measured 36 of 36 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-05T14:41:00.211355Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-08-05T14:41:01.099787Z

Reference resolution

35 of 35 outbound references displayed

  • verified exact0
  • verified fuzzy1
  • unresolved34
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation fbf2e421-312d-4335-9823-f4a50e5a4cb2 · outbound

This paper cites Phi-3 Technical Report: A Highly Capable Language Model Locally on Your Phone.

WebUIBench: A Comprehensive Benchmark for Evaluating Multimodal Large Language Models in WebUI-to-Code Phi-3 Technical Report: A Highly Capable Language Model Locally on Your Phone

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-07T05:27:50.039856Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:27:50.039856Z digest=sha256:d541de1febb04af6e227e9f6c67c0a7648004609a2bd65e3ff64034272a02542

Observation 671bea65-816c-4360-89b8-2fa35faaf777 · outbound

This paper cites The claude 3 model family: Opus, sonnet, haiku.

WebUIBench: A Comprehensive Benchmark for Evaluating Multimodal Large Language Models in WebUI-to-Code The claude 3 model family: Opus, sonnet, haiku

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-07T05:27:50.045581Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:27:50.045581Z digest=sha256:2de71b22700412750b1137fab9e5a4958281af9f9acfb45542c23cbf01bc5719

Observation 7492d6bf-6c4b-476e-9406-bd7068abfb66 · outbound

This paper cites an unresolved cited work.

WebUIBench: A Comprehensive Benchmark for Evaluating Multimodal Large Language Models in WebUI-to-Code Unresolved cited work

Reference 4

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:27:50.902427Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T05:27:50.050471Z digest=sha256:5b9c1527d96289f603df1ce0188a158c4bb11425214a36380d1078b75dcc4f35

Observation cf41afb1-0f8d-48df-9a40-536cffaa320b · outbound

This paper cites an unresolved cited work.

WebUIBench: A Comprehensive Benchmark for Evaluating Multimodal Large Language Models in WebUI-to-Code Unresolved cited work

Reference 5

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:27:50.889071Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T05:27:50.055521Z digest=sha256:badc28de2235e779a3f8330051b7710aa730368d3ffa898716e81155def42e5f

Observation 28a1da4d-8a62-400f-b4f3-171e136d64d8 · outbound

This paper cites A Survey on Evaluating Large Language Models in Code Generation Tasks.

WebUIBench: A Comprehensive Benchmark for Evaluating Multimodal Large Language Models in WebUI-to-Code A Survey on Evaluating Large Language Models in Code Generation Tasks

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-07T05:27:50.065829Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:27:50.065829Z digest=sha256:a5e1e7ffc74997a92f8364fc9d822ea78095926d7e91d59a86bb6b27c0f636a2

Observation 9f46b3ea-bc18-44f4-b119-946e3cd0bc83 · outbound

This paper cites Expanding Performance Boundaries of Open-Source Multimodal Models with Model, Data, and Test-Time Scaling.

WebUIBench: A Comprehensive Benchmark for Evaluating Multimodal Large Language Models in WebUI-to-Code Expanding Performance Boundaries of Open-Source Multimodal Models with Model, Data, and Test-Time Scaling

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-07T05:27:50.070819Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:27:50.070819Z digest=sha256:f058bb82c92e4d903546589cc622ca690c0e2e0daede4e61e47dfd9d52b45111

Observation b921ebde-1312-446b-a08d-498f458ec84a · outbound

This paper cites an unresolved cited work.

WebUIBench: A Comprehensive Benchmark for Evaluating Multimodal Large Language Models in WebUI-to-Code Unresolved cited work

Reference 8

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:27:50.859503Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T05:27:50.075882Z digest=sha256:c04d4716512e08bd4f7721d781b9c6f52c487d6efc8fab251b2f1acb0ccbc4b8

Observation cfd889bf-5ddb-45cc-9ddc-16eba5d4e6d9 · outbound

This paper cites an unresolved cited work.

WebUIBench: A Comprehensive Benchmark for Evaluating Multimodal Large Language Models in WebUI-to-Code Unresolved cited work

Reference 9

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:27:50.831773Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T05:27:50.085957Z digest=sha256:09a3bb122aa5533265359eac58c302bf50c9e62de36d32db1f1876e537b56384

Observation 67c5aa67-548f-46f2-8aab-7a1abb1231e9 · outbound

This paper cites ChatGLM: A Family of Large Language Models from GLM-130B to GLM-4 All Tools.

WebUIBench: A Comprehensive Benchmark for Evaluating Multimodal Large Language Models in WebUI-to-Code ChatGLM: A Family of Large Language Models from GLM-130B to GLM-4 All Tools

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-07T05:27:50.090673Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:27:50.090673Z digest=sha256:9a6f027da5b426837b5aca2eba81c3235a64f696c48aaa99c4f153054a6cd039

Observation c09c10b2-0920-4810-bcaa-59cd4a2b1532 · outbound

This paper cites an unresolved cited work.

WebUIBench: A Comprehensive Benchmark for Evaluating Multimodal Large Language Models in WebUI-to-Code Unresolved cited work

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-07T05:27:50.096470Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:27:50.096470Z digest=sha256:826957bbd4d20228e60f968c4d27826a69baf8b6c6970c0d01fcb344cc82d1ac

Observation 0412eb86-c49e-4a62-a859-f6b2f8b31154 · outbound

This paper cites GPT-4o System Card.

WebUIBench: A Comprehensive Benchmark for Evaluating Multimodal Large Language Models in WebUI-to-Code GPT-4o System Card

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-07T05:27:50.101164Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:27:50.101164Z digest=sha256:124a2ab950d773124ffc4b63f160da0c502284f0c4f0c5f7461cf8762eabc42f

Observation edbb6bfc-40fb-4cad-b17c-263868b35414 · outbound

This paper cites an unresolved cited work.

WebUIBench: A Comprehensive Benchmark for Evaluating Multimodal Large Language Models in WebUI-to-Code Unresolved cited work

Reference 13

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:27:50.817435Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T05:27:50.106403Z digest=sha256:11b50eabfb2526449eb27f04ec4fac15e5f844a07945518cc3306af961a84e1c

Observation 2ccfc924-d46b-43e4-90b0-abe1c759b149 · outbound

This paper cites an unresolved cited work.

WebUIBench: A Comprehensive Benchmark for Evaluating Multimodal Large Language Models in WebUI-to-Code Unresolved cited work

Reference 14

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:27:50.802724Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T05:27:50.116451Z digest=sha256:d43bc1a58754cd38143ac99725daaa5d051654ff46bfd62285db2cac4c65607c

Observation 38bb9698-9827-4001-bcca-7e9a2deee6d4 · outbound

This paper cites Unlocking the conversion of Web Screenshots into HTML Code with the WebSight Dataset.

WebUIBench: A Comprehensive Benchmark for Evaluating Multimodal Large Language Models in WebUI-to-Code Unlocking the conversion of Web Screenshots into HTML Code with the WebSight Dataset

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-07T05:27:50.111566Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:27:50.111566Z digest=sha256:4f3b9c4ce080d170b2a2ca75db87a5603c74e1f62c3448765e3083fa00bc4b10

Observation dd8e6060-44c9-4360-acc8-2401edbc7d62 · outbound

This paper cites A Survey on Multimodal Benchmarks: In the Era of Large AI Models.

WebUIBench: A Comprehensive Benchmark for Evaluating Multimodal Large Language Models in WebUI-to-Code A Survey on Multimodal Benchmarks: In the Era of Large AI Models

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-07T05:27:50.126766Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:27:50.126766Z digest=sha256:8d8b966317f8454e6b14162bd9e3c31d2a836e1582b32f8311a5ea88b7cbaf99

Observation 8eabf21d-8398-48fc-a512-5e5243d3ca70 · outbound

This paper cites SEED-Bench: Benchmarking Multimodal LLMs with Generative Comprehension.

WebUIBench: A Comprehensive Benchmark for Evaluating Multimodal Large Language Models in WebUI-to-Code SEED-Bench: Benchmarking Multimodal LLMs with Generative Comprehension

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-07T05:27:50.121929Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:27:50.121929Z digest=sha256:8b79ea4cedc24c5c963a60157bc7aba3c0b943045c3d29ae88f98e6b96ed8040

Observation 1f4d212c-0e72-4016-a0dd-0c62c367eb1e · outbound

This paper cites Ovis: Structural Embedding Alignment for Multimodal Large Language Model.

WebUIBench: A Comprehensive Benchmark for Evaluating Multimodal Large Language Models in WebUI-to-Code Ovis: Structural Embedding Alignment for Multimodal Large Language Model

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-07T05:27:50.136340Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:27:50.136340Z digest=sha256:58db9da069ff3acf31920ce33f2f924d00b89f690d673d28a212e81f7c9037f1

Observation 582d2760-34c3-4c13-a86b-8402941f4c61 · outbound

This paper cites an unresolved cited work.

WebUIBench: A Comprehensive Benchmark for Evaluating Multimodal Large Language Models in WebUI-to-Code Unresolved cited work

Reference 19

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:27:50.786294Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T05:27:50.131597Z digest=sha256:2bdd544e4180705d90b579875665256b87221d49e8b1bc827e5981891fa0419d

Observation 5de6a9c9-ac53-43b4-9682-0a7b347e2c0d · outbound

This paper cites an unresolved cited work.

WebUIBench: A Comprehensive Benchmark for Evaluating Multimodal Large Language Models in WebUI-to-Code Unresolved cited work

Reference 20

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:27:50.756555Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T05:27:50.145937Z digest=sha256:3faca187e436002d05b37e7f4d56480badc9ee6809a91f9de5e8bb58887bcacc

Observation e4eea27e-3d9f-48db-863d-a2a5b6a3883f · outbound

This paper cites an unresolved cited work.

WebUIBench: A Comprehensive Benchmark for Evaluating Multimodal Large Language Models in WebUI-to-Code Unresolved cited work

Reference 21

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:27:50.771983Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T05:27:50.140917Z digest=sha256:ddf760b686d26ef34393c2146758bf5ce440fba4dc75dd19d9a4289fdf33bb86

Observation d38f8e9b-6e87-41aa-a706-e357500e306c · outbound

This paper cites pix2code: Generating Code from a Graphical User Interface Screenshot.

WebUIBench: A Comprehensive Benchmark for Evaluating Multimodal Large Language Models in WebUI-to-Code pix2code: Generating Code from a Graphical User Interface Screenshot

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-07T05:27:50.155023Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:27:50.155023Z digest=sha256:de83cf2873fa727c5d6fe2a6e3c55791cdfc5abb3cd62f3da455e0b6368dbc70

Observation 28525e44-676a-45cc-af5e-c69bddc1ec7a · outbound

This paper cites an unresolved cited work.

WebUIBench: A Comprehensive Benchmark for Evaluating Multimodal Large Language Models in WebUI-to-Code Unresolved cited work

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-07T05:27:50.150390Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:27:50.150390Z digest=sha256:ee59698e1eee3aa8fb75c749ae8768dba7c7cf771a7426d5845965817ecaa11e

Observation 9ab7e1df-22c4-438f-9bc2-65e36a1a914d · outbound

This paper cites an unresolved cited work.

WebUIBench: A Comprehensive Benchmark for Evaluating Multimodal Large Language Models in WebUI-to-Code Unresolved cited work

Reference 24

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:27:50.724914Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T05:27:50.164789Z digest=sha256:64d5a8cd63119c7cc5eb370733995ab24e52113e9ac64ea764a854aa74446ccf

Observation 837eb6ba-c141-4f56-ba8c-a767dc3d279b · outbound

This paper cites an unresolved cited work.

WebUIBench: A Comprehensive Benchmark for Evaluating Multimodal Large Language Models in WebUI-to-Code Unresolved cited work

Reference 25

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:27:50.739755Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T05:27:50.159978Z digest=sha256:422e9549c061cd281c0c6b99c0ef850e8c6964006d4fe01b3db83bdc14015161

Observation 0052eaaa-fb5e-4114-977c-0ab8c2a20266 · outbound

This paper cites an unresolved cited work.

WebUIBench: A Comprehensive Benchmark for Evaluating Multimodal Large Language Models in WebUI-to-Code Unresolved cited work

Reference 26

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:27:50.710121Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T05:27:50.175237Z digest=sha256:c666572caa45c0c8f1da1915da974f5c8f1b4442c435452118022987c462f964

Observation df3b7bd5-453f-47de-beef-0f6e6d871a74 · outbound

This paper cites Design2Code: Benchmarking Multimodal Code Generation for Automated Front-End Engineering.

WebUIBench: A Comprehensive Benchmark for Evaluating Multimodal Large Language Models in WebUI-to-Code Design2Code: Benchmarking Multimodal Code Generation for Automated Front-End Engineering

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-07T05:27:50.170369Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:27:50.170369Z digest=sha256:8a2c048398252e2d3374e04b1a6c647a7bd87f6213da8a6f8d28407e285fcbcc

Observation 5f61bfa5-2bd9-43bd-8326-0db51f8349ee · outbound

This paper cites an unresolved cited work.

WebUIBench: A Comprehensive Benchmark for Evaluating Multimodal Large Language Models in WebUI-to-Code Unresolved cited work

Reference 28

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:27:50.695455Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T05:27:50.189437Z digest=sha256:f438e2b237a9cf56451e86cce54854d0e368f679db2590e658206b11b9f376f5

Observation 57a8284c-cc14-48a0-8c93-e485495524ae · outbound

This paper cites Gemini 1.5: Unlocking multimodal understanding across millions of tokens of context.

WebUIBench: A Comprehensive Benchmark for Evaluating Multimodal Large Language Models in WebUI-to-Code Gemini 1.5: Unlocking multimodal understanding across millions of tokens of context

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-07T05:27:50.180028Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:27:50.180028Z digest=sha256:9f1d3e1342e8680795c3ce8b258aa156975da582d548108dc741657aedbe83c5

Observation 5f0e6284-b7ec-4559-85ce-9e1922115354 · outbound

This paper cites Qwen2-VL: Enhancing Vision-Language Model's Perception of the World at Any Resolution.

WebUIBench: A Comprehensive Benchmark for Evaluating Multimodal Large Language Models in WebUI-to-Code Qwen2-VL: Enhancing Vision-Language Model's Perception of the World at Any Resolution

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-07T05:27:50.184733Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:27:50.184733Z digest=sha256:35292e0f8ce47a19895d64f66d11a0a084519520014e660edc86258bcf73c169

Observation a77a577e-0620-4f7e-a949-71cbfd378f93 · outbound

This paper cites an unresolved cited work.

WebUIBench: A Comprehensive Benchmark for Evaluating Multimodal Large Language Models in WebUI-to-Code Unresolved cited work

Reference 31

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:27:50.680815Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T05:27:50.202973Z digest=sha256:679900b24170d52a1880c1ac2d307d4ea6c652a197e117d24b2dc82dfa8eee60

Observation 3a0af785-abc6-4c3d-9541-a5d56e14c59e · outbound

This paper cites MiniCPM-V: A GPT-4V Level MLLM on Your Phone.

WebUIBench: A Comprehensive Benchmark for Evaluating Multimodal Large Language Models in WebUI-to-Code MiniCPM-V: A GPT-4V Level MLLM on Your Phone

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-07T05:27:50.194019Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:27:50.194019Z digest=sha256:15dec44123c0a02c578a464db94858c070099e20f06382074e3d96f6837d8e9d

Observation ad472f39-e63f-48cc-9896-f8f7cb521f2d · outbound

This paper cites Yi: Open Foundation Models by 01.AI.

WebUIBench: A Comprehensive Benchmark for Evaluating Multimodal Large Language Models in WebUI-to-Code Yi: Open Foundation Models by 01.AI

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-07T05:27:50.198375Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:27:50.198375Z digest=sha256:fe64e030a3489a62b692f02a1288e7dd4fae8f4a1313a26fc60077e1d833e1fb

Observation 541a29be-d111-4864-844e-4c3dd74fd1d9 · outbound

This paper cites Web2Code: A Large-scale Webpage-to-Code Dataset and Evaluation Framework for Multimodal LLMs.

WebUIBench: A Comprehensive Benchmark for Evaluating Multimodal Large Language Models in WebUI-to-Code Web2Code: A Large-scale Webpage-to-Code Dataset and Evaluation Framework for Multimodal LLMs

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-07T05:27:50.207219Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:27:50.207219Z digest=sha256:fdd8cc1979edd85793dc1d0a676b4bb1b8f472e4869e2890e81279fdb4907906

Observation 5e3e09cf-24ad-47ad-bca2-f61ddcdaa6f9 · outbound

This paper cites an unresolved cited work.

WebUIBench: A Comprehensive Benchmark for Evaluating Multimodal Large Language Models in WebUI-to-Code Unresolved cited work

Reference 2005

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:27:50.874432Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T05:27:50.060579Z digest=sha256:eb0387d8fd27f495caa72b460e2f05bccc32860d8ec97e4ae7acbdd3674553c7

Observation 09fb2401-0a1d-49fc-b72f-a2017d08e018 · outbound

This paper cites arXiv preprint.

WebUIBench: A Comprehensive Benchmark for Evaluating Multimodal Large Language Models in WebUI-to-Code arXiv preprint

Reference 2024

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:27:50.845803Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T05:27:50.080768Z digest=sha256:2e18bcf7eee5dcac9e913a958c37e6772ef7cf7dc35cdb3b941753a7425cc9d6

Pith citing papers

Observation 9aede403-8318-4a05-a146-1b6eace7ca0d · inbound

Pattern over Pixels: Measuring Pattern Completion Bias in Multimodal Code Generation cites this paper.

Pattern over Pixels: Measuring Pattern Completion Bias in Multimodal Code Generation WebUIBench: A Comprehensive Benchmark for Evaluating Multimodal Large Language Models in WebUI-to-Code

Reference 23

Resolution
verified exact
local_arxiv, observed 2026-08-05T14:41:01.107702Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-05T14:41:00.211355Z digest=sha256:862cc71ddce45fd98d1943158bda5bfb18a5d6f783c4b75b4f992494ac93d6e6