Pith. sign in

Paper Citation Record · LEDGER

AesRM: Improving Video Aesthetics with Expert-Level Feedback

As of 3 August 2026, this Paper Citation Record lists 15 of 15 outbound references and 2 inbound Pith citation observations for arXiv:2604.28078.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2604.28078 v1

Coverage vector

measured 15 of 15 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-07-11T11:50:26.030339Z

measured 17 of 17 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-03T06:30:56.289259+00:00

measured 2 of 2 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-06-30T18:00:39.556424Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-06-30T18:04:58.053921Z

Reference resolution

15 of 15 outbound references displayed

  • verified exact0
  • verified fuzzy3
  • unresolved6
  • parse uncertain0
  • malformed identifier1
  • metadata mismatch5

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 55b41d5c-1d02-4242-9ac1-22aa495326de · outbound

This paper cites Improving Dynamic Object Interactions in Text-to-Video Generation with AI Feedback.

AesRM: Improving Video Aesthetics with Expert-Level Feedback Improving Dynamic Object Interactions in Text-to-Video Generation with AI Feedback

Reference 1

Resolution
metadata mismatch
local_arxiv, observed 2026-05-09T05:05:13.617198Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.

source=pdf_text observed=2026-07-11T11:50:26.030339Z digest=sha256:ead190ac1ef3e3254bbe9f3f21921c3d473e224296a3a62ffbbc3616e8473ba2

Observation cbef0ede-e0ec-4ff6-b7a4-0b45d3c81446 · outbound

This paper cites GPT-4 Technical Report.

AesRM: Improving Video Aesthetics with Expert-Level Feedback GPT-4 Technical Report

Reference 2

Resolution
metadata mismatch
local_arxiv, observed 2026-05-12T10:31:30.650227Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.

source=pdf_text observed=2026-05-07T05:42:03.688350Z digest=sha256:dfe89b939d790bf605c58a600a61317ea97973a1ce448b059bfe6e2a841fff0f

Observation 9e675bf4-c75d-4fb6-b81b-40cb33d31a90 · outbound

This paper cites Wan: Open and Advanced Large-Scale Video Generative Models.

AesRM: Improving Video Aesthetics with Expert-Level Feedback Wan: Open and Advanced Large-Scale Video Generative Models

Reference 3

Resolution
metadata mismatch
local_arxiv, observed 2026-05-12T10:31:30.655756Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.

source=pdf_text observed=2026-05-07T05:42:03.688350Z digest=sha256:fa1c213a063b94be5ee13aa25dbb3e09f5f60c0f0e3115c0f07619e13c3e9449

Observation 94c97aaf-bf80-4fdf-9ce0-9ebb3e7c264a · outbound

This paper cites DanceGRPO: Unleashing GRPO on Visual Generation.

AesRM: Improving Video Aesthetics with Expert-Level Feedback DanceGRPO: Unleashing GRPO on Visual Generation

Reference 4

Resolution
metadata mismatch
local_arxiv, observed 2026-05-12T10:36:28.613267Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.

source=pdf_text observed=2026-05-07T05:42:03.688350Z digest=sha256:25a526272f00122be140d885bde6e315feecfac54c0e7febe3ea6acaa75e0f51

Observation f1fc2eaf-e21a-4372-bd60-1af5bded01c1 · outbound

This paper cites InternVL3: Exploring Advanced Training and Test-Time Recipes for Open-Source Multimodal Models.

AesRM: Improving Video Aesthetics with Expert-Level Feedback InternVL3: Exploring Advanced Training and Test-Time Recipes for Open-Source Multimodal Models

Reference 5

Resolution
metadata mismatch
local_arxiv, observed 2026-05-12T10:31:30.661239Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.

source=pdf_text observed=2026-05-07T05:42:03.688350Z digest=sha256:7dc376bdf57e762226d8f9ed84cc08f4a1f0a317db79850f9d0a89f2e23f4c9f

Observation efc9ea92-24a9-4507-995d-963ed4f52dfd · outbound

This paper cites an unresolved cited work.

AesRM: Improving Video Aesthetics with Expert-Level Feedback Unresolved cited work

Reference 6

Resolution
unresolved
raw_fallback, observed 2026-05-27T10:33:59.249091Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.

source=pdf_text observed=2026-05-07T05:42:03.688350Z digest=sha256:3e968694b167e694915e974d303ee54cd0f9e390546e8b6fc3a220e5d1a04537

Observation abf03cca-e2bf-489a-9f7e-91e58043f210 · outbound

This paper cites an unresolved cited work.

AesRM: Improving Video Aesthetics with Expert-Level Feedback Unresolved cited work

Reference 7

Resolution
unresolved
raw_fallback, observed 2026-05-27T10:33:59.229010Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.

source=pdf_text observed=2026-05-07T05:42:03.688350Z digest=sha256:6a0d8143ee8a2993f344b30615b69edf5a7059ab0384d10a440712b89412fc77

Observation 77493d8e-81c4-40ba-8b13-404b0832129a · outbound

This paper cites an unresolved cited work.

AesRM: Improving Video Aesthetics with Expert-Level Feedback Unresolved cited work

Reference 8

Resolution
unresolved
raw_fallback, observed 2026-05-27T10:33:59.241399Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.

source=pdf_text observed=2026-05-07T05:42:03.688350Z digest=sha256:aa63ef109856ac60be57beeb18978531d49ee893f894bd0444fda7c91607e04f

Observation 29436fb8-e60f-4bf6-9663-120fb3dbbeee · outbound

This paper cites Light Style.

AesRM: Improving Video Aesthetics with Expert-Level Feedback Light Style

Reference 9

Resolution
malformed identifier
raw_fallback, observed 2026-05-27T10:33:59.237193Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.

source=pdf_text observed=2026-05-07T05:42:03.688350Z digest=sha256:a4afa86639c0f2d854ad3854ff25d7c6effee697b798cf25b7ff1e11c1ed104f

Observation 9c364800-1680-4876-b7b1-1ba02b646ea8 · outbound

This paper cites an unresolved cited work.

AesRM: Improving Video Aesthetics with Expert-Level Feedback Unresolved cited work

Reference 10

Resolution
unresolved
raw_fallback, observed 2026-05-27T10:33:59.233547Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.

source=pdf_text observed=2026-05-07T05:42:03.688350Z digest=sha256:b0e7bd4a0b67e9b56a2906ff61f6ce96d279a7eccba280ee8421601d9a55d16a

Observation 3f02a3bc-bd66-4760-9aae-05169ab3e327 · outbound

This paper cites For example: Video A underperforms Video B in visual aesthetics, while the two are comparable in visual fidelity and visual plausibility.

AesRM: Improving Video Aesthetics with Expert-Level Feedback For example: Video A underperforms Video B in visual aesthetics, while the two are comparable in visual fidelity and visual plausibility

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-05-27T10:33:59.245509Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.

source=pdf_text observed=2026-05-07T05:42:03.688350Z digest=sha256:d5c3a25ceb5bcdaabbbd7dafe1caebb4e2d7c8574fa608217aaefbfca9f24e76

Observation 13041c59-bbee-42de-8a5e-db9b6a804e64 · outbound

This paper cites Table 9: Full system prompt for AesRM-CoT.

AesRM: Improving Video Aesthetics with Expert-Level Feedback Table 9: Full system prompt for AesRM-CoT

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-05-27T10:33:59.225055Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.

source=pdf_text observed=2026-05-07T05:42:03.688350Z digest=sha256:884e43c50cec5a4ac95fda0ca77254188b1267efeef9c0e35cce31aea00a9827

Observation 04d36ffe-70eb-44be-aaf7-5ecd1f68377a · outbound

This paper cites an unresolved cited work.

AesRM: Improving Video Aesthetics with Expert-Level Feedback Unresolved cited work

Reference 13

Resolution
unresolved
raw_fallback, observed 2026-05-27T10:33:59.255739Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.

source=pdf_text observed=2026-05-07T05:42:03.688350Z digest=sha256:f2b39508ac23af23387485c9aa2d81f1e4872167d2f57c843e1a376f4d1e897c

Observation b5116620-74c8-4d58-8c27-726ae4c013ad · outbound

This paper cites an unresolved cited work.

AesRM: Improving Video Aesthetics with Expert-Level Feedback Unresolved cited work

Reference 14

Resolution
unresolved
raw_fallback, observed 2026-05-27T10:33:59.252750Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.

source=pdf_text observed=2026-05-07T05:42:03.688350Z digest=sha256:ef51a4741b6da235aa2e467ee9ca4cbafa1eab696245d4f2965534872594dc40

Observation 7aa91d4d-b330-4471-be34-905b4a570640 · outbound

This paper cites $! & % #.

AesRM: Improving Video Aesthetics with Expert-Level Feedback $! & % #

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-05-27T10:33:59.221095Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.

source=pdf_text observed=2026-05-07T05:42:03.688350Z digest=sha256:57461fa3a759d8b66254ed1fe23e883f75c5e40aa8597622e9c805da87bac9f3

Pith citing papers

Observation 336ca822-3d25-4145-8f8e-12c4bc52fb9a · inbound

MSAVBench: Towards Comprehensive and Reliable Evaluation of Multi-Shot Audio-Video Generation cites this paper.

MSAVBench: Towards Comprehensive and Reliable Evaluation of Multi-Shot Audio-Video Generation AesRM: Improving Video Aesthetics with Expert-Level Feedback

Reference 21

Resolution
verified exact
local_arxiv, observed 2026-05-20T05:18:03.056737Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.

source=pdf_text observed=2026-05-20T05:17:11.690484Z digest=sha256:eda540107fae5520e446aba1677acb20d0a75258834394d3d94d4c76c1eb1b74

Observation 073be41a-85d1-4974-833e-5255aec4d323 · inbound

MSAVBench: Towards Comprehensive and Reliable Evaluation of Multi-Shot Audio-Video Generation cites this paper.

MSAVBench: Towards Comprehensive and Reliable Evaluation of Multi-Shot Audio-Video Generation AesRM: Improving Video Aesthetics with Expert-Level Feedback

Reference 21

Resolution
verified exact
local_arxiv, observed 2026-06-30T18:04:58.055663Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.

source=pdf_text observed=2026-06-30T18:00:39.556424Z digest=sha256:6925f04b92279f224f943283401165d6183d7e00f2103f9900de875e6e93bda2