Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-11T13:10:34.717993Z
Paper Citation Record · LEDGER
As of 14 August 2026, this Paper Citation Record lists 75 of 75 outbound references and 0 inbound Pith citation observations for arXiv:2608.09666.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-11T13:10:34.717993Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-14T06:32:32.682623+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links
A source-named dated measurement, never combined with another source.
Source: cited_works
75 of 75 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 0ce8ec41-0033-4d84-ae98-2b7e0883dc5b · outbound
Open Evaluation Agent: Efficient and Promptable Evaluation of Visual Generative Models Qwen2.5 Technical Report
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 67d9b3e1-a1aa-46af-969c-78eef55aa7fd · outbound
Open Evaluation Agent: Efficient and Promptable Evaluation of Visual Generative Models Evaluation Agent: Efficient and promptable evaluation framework for visual generative models,
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation d51bb55d-ef9d-4530-b004-e4ac10c922c1 · outbound
Open Evaluation Agent: Efficient and Promptable Evaluation of Visual Generative Models Towards Accurate Generative Models of Video: A New Metric & Challenges
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a948c38c-dd87-4d02-a0ee-1618637cd98d · outbound
Open Evaluation Agent: Efficient and Promptable Evaluation of Visual Generative Models Gans trained by a two time-scale update rule converge to a local nash equilibrium,
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8683498f-8e6a-449c-af39-1306ee3b6cb7 · outbound
Open Evaluation Agent: Efficient and Promptable Evaluation of Visual Generative Models T2i-compbench: A comprehensive benchmark for open-world compositional text-to-image generation,
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 56ace4f0-c158-4dda-9b57-846df8fd8610 · outbound
Open Evaluation Agent: Efficient and Promptable Evaluation of Visual Generative Models VBench: Comprehensive benchmark suite for video generative models,
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1926fd13-2cc2-455f-8a17-e65e72cd35a9 · outbound
Open Evaluation Agent: Efficient and Promptable Evaluation of Visual Generative Models Denoising diffusion probabilistic models,
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation cb7107c7-370c-4dba-8432-49fb0ed245a0 · outbound
Open Evaluation Agent: Efficient and Promptable Evaluation of Visual Generative Models Frozen in time: A joint video and image encoder for end-to-end retrieval,
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation aa6da551-973d-4fe0-bc41-6a9f6bc117a2 · outbound
Open Evaluation Agent: Efficient and Promptable Evaluation of Visual Generative Models Panda- 70m: Captioning 70m videos with multiple cross-modality teachers,
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation bfe0e96f-3486-4fc1-9e1e-b1765795bc47 · outbound
Open Evaluation Agent: Efficient and Promptable Evaluation of Visual Generative Models Video enhancement with task-oriented flow,
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation e3b501df-378b-46d2-9218-9ca7e8f17a1b · outbound
Open Evaluation Agent: Efficient and Promptable Evaluation of Visual Generative Models Vbench++: Comprehensive and versatile benchmark suite for video generative models,
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 466618b1-645e-4339-a2d3-f6335af20e77 · outbound
Open Evaluation Agent: Efficient and Promptable Evaluation of Visual Generative Models Evalcrafter: Benchmarking and evaluating large video generation models,
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0b97338e-38c6-4012-abba-3d1f69e9101a · outbound
Open Evaluation Agent: Efficient and Promptable Evaluation of Visual Generative Models Llamafactory: Unified efficient fine-tuning of 100+ language models,
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation e14bbe5b-0663-4032-a70d-f4411876dc4e · outbound
Open Evaluation Agent: Efficient and Promptable Evaluation of Visual Generative Models High- resolution image synthesis with latent diffusion models,
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4692b390-534f-43ac-90d2-916e6ee2979d · outbound
Open Evaluation Agent: Efficient and Promptable Evaluation of Visual Generative Models Sdxl: Improving latent diffusion models for high-resolution image synthesis,
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2b9bf8ef-553d-4166-b1b1-586c0d30964c · outbound
Open Evaluation Agent: Efficient and Promptable Evaluation of Visual Generative Models Scaling rectified flow transformers for high-resolution image synthesis,
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 891eab10-9724-4dbc-9f45-efea4e269154 · outbound
Open Evaluation Agent: Efficient and Promptable Evaluation of Visual Generative Models Latte: Latent diffusion transformer for video generation,
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 32131614-b96d-4de5-9171-82b66ac2b60f · outbound
Open Evaluation Agent: Efficient and Promptable Evaluation of Visual Generative Models ModelScope Text-to-Video Technical Report
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a7703b7b-ae9a-4d25-bdf8-919011621982 · outbound
Open Evaluation Agent: Efficient and Promptable Evaluation of Visual Generative Models Videocrafter2: Overcoming data limitations for high-quality video diffu- sion models,
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation cb222d48-952f-41e6-af15-0d0ed195691b · outbound
Open Evaluation Agent: Efficient and Promptable Evaluation of Visual Generative Models Latent Video Diffusion Models for High-Fidelity Long Video Generation
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3c65ca84-5049-493e-b634-5488567a0402 · outbound
Open Evaluation Agent: Efficient and Promptable Evaluation of Visual Generative Models Lavie: High-quality video generation with cascaded latent diffusion models,
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9ef1105c-3b0d-4988-9341-7a894ac0c23b · outbound
Open Evaluation Agent: Efficient and Promptable Evaluation of Visual Generative Models VBench-2.0: Advancing Video Generation Benchmark Suite for Intrinsic Faithfulness
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e2a1f3a3-abfd-4173-8acd-7a4248f1b9c0 · outbound
Open Evaluation Agent: Efficient and Promptable Evaluation of Visual Generative Models Holistic evaluation of text- to-image models,
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 1cb80ab0-f00c-4da4-be17-8c706ec867a1 · outbound
Open Evaluation Agent: Efficient and Promptable Evaluation of Visual Generative Models T2v- compbench: A comprehensive benchmark for compositional text-to-video generation,
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 1980a47c-d758-4e6c-a947-3e4be4d134ee · outbound
Open Evaluation Agent: Efficient and Promptable Evaluation of Visual Generative Models RealUnify: Do unified models truly benefit from unification? A comprehensive benchmark,
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 1a877a63-62b2-4459-ab31-4171aa424c08 · outbound
Open Evaluation Agent: Efficient and Promptable Evaluation of Visual Generative Models Uni-MMMU: A massive multi-discipline multimodal unified benchmark,
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation a76b7fe1-b4a2-4447-b7a3-44b6ba3eae7e · outbound
Open Evaluation Agent: Efficient and Promptable Evaluation of Visual Generative Models Unison: Benchmarking unified multimodal models via synergistic understanding and generation,
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation bb6af445-5f93-4967-a2db-20498bdaca42 · outbound
Open Evaluation Agent: Efficient and Promptable Evaluation of Visual Generative Models Multi-dimensional evaluation of text summarization with in-context learning,
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f360cad3-a1a7-4b5c-820d-9cabe5ff4e23 · outbound
Open Evaluation Agent: Efficient and Promptable Evaluation of Visual Generative Models Can Large Language Models Be an Alternative to Human Evaluations?
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5138a724-6483-4c4a-8de8-b6b274f43bbf · outbound
Open Evaluation Agent: Efficient and Promptable Evaluation of Visual Generative Models Gptscore: Evaluate as you desire,
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 5f82c26d-891f-4dca-9760-8b0e6f1bffef · outbound
Open Evaluation Agent: Efficient and Promptable Evaluation of Visual Generative Models Judging llm-as-a-judge with mt-bench and chatbot arena,
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 27b06c18-8f3c-4548-affb-c52c9f8ec0e5 · outbound
Open Evaluation Agent: Efficient and Promptable Evaluation of Visual Generative Models Exploring the reliability of large language models as customized evaluators for diverse nlp tasks,
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 3cfd6982-883c-45c2-a497-0c5ae3031012 · outbound
Open Evaluation Agent: Efficient and Promptable Evaluation of Visual Generative Models Au- tonomous evaluation and refinement of digital agents,
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation d9a95d08-7651-4dd2-8b5f-8e93d92bbc5f · outbound
Open Evaluation Agent: Efficient and Promptable Evaluation of Visual Generative Models A survey on large language model based autonomous agents,
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 317db75e-6f79-4bc0-9b4a-c7b907d44b51 · outbound
Open Evaluation Agent: Efficient and Promptable Evaluation of Visual Generative Models From generation to judgment: Opportunities and challenges of LLM-as-a-judge,
Reference 35
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 23bfe681-a927-4fd4-b297-1f8a6d9c75c4 · outbound
Open Evaluation Agent: Efficient and Promptable Evaluation of Visual Generative Models Agent-as-a-Judge,
Reference 36
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 97456063-a9d5-4723-93d8-aa98e61173d0 · outbound
Open Evaluation Agent: Efficient and Promptable Evaluation of Visual Generative Models Agent-as-a-Judge: Evaluate agents with agents,
Reference 37
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 6cfbbefe-85e5-40ea-b8f4-66c4e846f40b · outbound
Open Evaluation Agent: Efficient and Promptable Evaluation of Visual Generative Models Adaptively profiling models with task elicitation,
Reference 38
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 069b477f-27f9-436d-ae62-8c70bda07030 · outbound
Open Evaluation Agent: Efficient and Promptable Evaluation of Visual Generative Models Mind2Web 2: Evaluating agentic search with Agent-as-a-Judge,
Reference 39
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 59f83f50-386a-4d89-b8ad-59be708d796a · outbound
Open Evaluation Agent: Efficient and Promptable Evaluation of Visual Generative Models One-Eval: An agentic system for automated and traceable LLM evaluation,
Reference 40
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cb060ea8-c4f9-45ba-a6ee-15d2b48e2974 · outbound
Open Evaluation Agent: Efficient and Promptable Evaluation of Visual Generative Models AgenticEval: Toward agentic and self-evolving safety evaluation of large language models,
Reference 41
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 1c1d5067-e92d-446a-b227-13c7d55f2d7b · outbound
Open Evaluation Agent: Efficient and Promptable Evaluation of Visual Generative Models Evaluating Hallucination in Text-to-Image Diffusion Models with Scene-Graph based Question-Answering Agent
Reference 42
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation ef08f582-8260-4a79-b72f-d8ce8c51fe6b · outbound
Open Evaluation Agent: Efficient and Promptable Evaluation of Visual Generative Models A unified agentic framework for evaluating conditional image generation,
Reference 43
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation f8f1d3d3-c7d6-4506-a6e1-fcb7cc07e9f9 · outbound
Open Evaluation Agent: Efficient and Promptable Evaluation of Visual Generative Models EdiVal-Agent: An object-centric framework for automated, fine- grained evaluation of multi-turn editing,
Reference 44
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2765068d-88ee-48b8-9d9d-6331df4fb633 · outbound
Open Evaluation Agent: Efficient and Promptable Evaluation of Visual Generative Models RewardHarness: Self-Evolving Agentic Post-Training
Reference 45
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 3a1ae944-f003-49cd-bb5a-9c262d9b31a4 · outbound
Open Evaluation Agent: Efficient and Promptable Evaluation of Visual Generative Models VideoGen-Eval: Agent-based System for Video Generation Evaluation
Reference 46
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a568ed33-ab06-4704-816f-5183ef1ddc14 · outbound
Open Evaluation Agent: Efficient and Promptable Evaluation of Visual Generative Models VQQA: An agentic approach for video evaluation and quality improvement,
Reference 47
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0a055ea0-48d5-4d28-8c40-4dbf0e375dd3 · outbound
Open Evaluation Agent: Efficient and Promptable Evaluation of Visual Generative Models VideoArgus: Agentic Rubric-Grounded Unified Evaluation for Video Generation and Editing
Reference 48
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation a4915c07-cd28-4da8-9f87-231455740a72 · outbound
Open Evaluation Agent: Efficient and Promptable Evaluation of Visual Generative Models AgentRewardBench: Evaluating automatic evaluations of web agent trajectories,
Reference 49
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation b805a161-23ef-42b2-899e-54b29dcf7edd · outbound
Open Evaluation Agent: Efficient and Promptable Evaluation of Visual Generative Models AJ-Bench: Benchmarking Agent-as-a-Judge for environment-aware evaluation,
Reference 50
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 55fbe9c5-8aad-44a1-9f4b-1665da83e741 · outbound
Open Evaluation Agent: Efficient and Promptable Evaluation of Visual Generative Models Time to REFLECT: Can We Trust LLM Judges for Evidence-based Research Agents?
Reference 51
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 4892a728-5e49-4714-986d-8ce6d38c1493 · outbound
Open Evaluation Agent: Efficient and Promptable Evaluation of Visual Generative Models Establishing best practices in building rigorous agentic benchmarks,
Reference 52
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation f5b72079-8318-4c8f-9af9-c07a9d4f51c0 · outbound
Open Evaluation Agent: Efficient and Promptable Evaluation of Visual Generative Models Webarena: A realistic web environment for building autonomous agents,
Reference 53
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 884f717c-c04c-4882-8921-73d07daa2807 · outbound
Open Evaluation Agent: Efficient and Promptable Evaluation of Visual Generative Models Osworld: Benchmarking multimodal agents for open-ended tasks in real computer environments,
Reference 54
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 5ec24db2-4335-42c0-8707-af4d624159d8 · outbound
Open Evaluation Agent: Efficient and Promptable Evaluation of Visual Generative Models Mobile-agent: Autonomous multi-modal mobile device agent with visual perception,
Reference 55
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 0779f335-e4bc-40aa-a9e8-7f7f13f1c724 · outbound
Open Evaluation Agent: Efficient and Promptable Evaluation of Visual Generative Models Omniact: A dataset and benchmark for enabling multimodal generalist autonomous agents for desktop and web,
Reference 56
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation d20b6cec-7b9a-4352-983f-8470dfa8f8c4 · outbound
Open Evaluation Agent: Efficient and Promptable Evaluation of Visual Generative Models MMInA: Benchmarking multihop multimodal internet agents,
Reference 57
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 15e09f98-67a5-4fc6-abea-43c5421f0cea · outbound
Open Evaluation Agent: Efficient and Promptable Evaluation of Visual Generative Models Chain-of-thought prompting elicits reasoning in large language models,
Reference 58
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 394a836c-5f31-404a-938a-26428e3f7bd7 · outbound
Open Evaluation Agent: Efficient and Promptable Evaluation of Visual Generative Models Large language models are zero-shot reasoners,
Reference 59
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 62dfbfd9-5c76-4a20-b42d-f4073384bb7c · outbound
Open Evaluation Agent: Efficient and Promptable Evaluation of Visual Generative Models React: Synergizing reasoning and acting in language models,
Reference 60
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation ec597b3a-aba7-4ed7-9f8a-611813990e26 · outbound
Open Evaluation Agent: Efficient and Promptable Evaluation of Visual Generative Models Tree of thoughts: Deliberate problem solving with large language models,
Reference 61
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e55cfbbf-f400-478b-991e-9584cd3f4c0e · outbound
Open Evaluation Agent: Efficient and Promptable Evaluation of Visual Generative Models Algorithm of thoughts: Enhancing exploration of ideas in large language models,
Reference 62
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation a892988b-bd30-44b1-9cfe-1025c985f153 · outbound
Open Evaluation Agent: Efficient and Promptable Evaluation of Visual Generative Models Graph of thoughts: Solving elaborate problems with large language models,
Reference 63
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 76b191dc-095c-4924-bd9b-b451039044fc · outbound
Open Evaluation Agent: Efficient and Promptable Evaluation of Visual Generative Models Self-consistency improves chain of thought reasoning in language models,
Reference 64
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation bb21bdbe-597c-4960-935f-d2c26ea82ddb · outbound
Open Evaluation Agent: Efficient and Promptable Evaluation of Visual Generative Models Reflexion: Language agents with verbal reinforcement learning,
Reference 65
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 57c7b42e-ed26-4ad6-927a-87324c483f98 · outbound
Open Evaluation Agent: Efficient and Promptable Evaluation of Visual Generative Models Api-bank: A comprehensive benchmark for tool-augmented llms,
Reference 66
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation c67a8f2d-6640-446c-acfe-2e2b536b9b3f · outbound
Open Evaluation Agent: Efficient and Promptable Evaluation of Visual Generative Models Toolllm: Facilitating large language models to master 16000+ real-world apis,
Reference 67
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation b15b39b5-b8e5-4919-81bd-aa13ab787c5d · outbound
Open Evaluation Agent: Efficient and Promptable Evaluation of Visual Generative Models Toolformer: Language models can teach themselves to use tools,
Reference 68
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8154be5d-4f83-4782-a13c-cd724398a4cc · outbound
Open Evaluation Agent: Efficient and Promptable Evaluation of Visual Generative Models Gorilla: Large language model connected with massive apis,
Reference 69
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation c8a17faa-7dd2-4f7f-a83c-43090825a05b · outbound
Open Evaluation Agent: Efficient and Promptable Evaluation of Visual Generative Models T2i-r1: Reinforcing image generation with collaborative semantic-level and token-level cot,
Reference 70
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation b61177a4-7c74-4244-b54e-69e585e58440 · outbound
Open Evaluation Agent: Efficient and Promptable Evaluation of Visual Generative Models VChain: Chain-of-visual-thought for reasoning in video generation,
Reference 71
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 8882e30d-a1af-4e9b-9776-c976ddb3d579 · outbound
Open Evaluation Agent: Efficient and Promptable Evaluation of Visual Generative Models $\tau^2$-Bench: Evaluating Conversational Agents in a Dual-Control Environment
Reference 72
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e7707f2b-9fd1-4c3b-acad-42cc919d19f1 · outbound
Open Evaluation Agent: Efficient and Promptable Evaluation of Visual Generative Models An illusion of progress? assessing the current state of web agents,
Reference 73
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation d77e4872-7f73-4950-94e3-8a347616e4d0 · outbound
Open Evaluation Agent: Efficient and Promptable Evaluation of Visual Generative Models PaperBench: Evaluating AI’s ability to replicate AI research,
Reference 74
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation c4903075-9b2a-4fab-a38f-4b73a59b0173 · outbound
Open Evaluation Agent: Efficient and Promptable Evaluation of Visual Generative Models Clipscore: A reference-free evaluation metric for image captioning,
Reference 75
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
No inbound Pith citation observations are available.