Pith. sign in

Paper Citation Record · LEDGER

AesRM: Improving Video Aesthetics with Expert-Level Feedback

As of 3 August 2026, this Paper Citation Record lists 15 of 15 outbound references and 2 inbound Pith citation observations for arXiv:2604.28078.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2604.28078 v1

Coverage vector

measured 15 of 15 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-07-11T11:50:26.030339Z

measured 17 of 17 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-02T06:30:47.504484+00:00

measured 2 of 2 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-06-30T18:00:39.556424Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-06-30T18:04:58.053921Z

Reference resolution

15 of 15 outbound references displayed

  • verified exact0
  • verified fuzzy3
  • unresolved6
  • parse uncertain0
  • malformed identifier1
  • metadata mismatch5

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 55b41d5c-1d02-4242-9ac1-22aa495326de · outbound

This paper cites Improving Dynamic Object Interactions in Text-to-Video Generation with AI Feedback.

AesRM: Improving Video Aesthetics with Expert-Level Feedback Improving Dynamic Object Interactions in Text-to-Video Generation with AI Feedback

Reference 1

Resolution
metadata mismatch
local_arxiv, observed 2026-05-09T05:05:13.617198Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-02T06:30:47.504484+00:00.

source=pdf_text observed=2026-07-11T11:50:26.030339Z digest=sha256:f57af68988fb6e2cd43f017698ea7bba57e7a3dca312931c23d43c30d4e45974

Observation cbef0ede-e0ec-4ff6-b7a4-0b45d3c81446 · outbound

This paper cites GPT-4 Technical Report.

AesRM: Improving Video Aesthetics with Expert-Level Feedback GPT-4 Technical Report

Reference 2

Resolution
metadata mismatch
local_arxiv, observed 2026-05-12T10:31:30.650227Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-02T06:30:47.504484+00:00.

source=pdf_text observed=2026-05-07T05:42:03.688350Z digest=sha256:a77169540877af196244593a2f0c080e28c2684fef43a680bc052216f053c759

Observation 9e675bf4-c75d-4fb6-b81b-40cb33d31a90 · outbound

This paper cites Wan: Open and Advanced Large-Scale Video Generative Models.

AesRM: Improving Video Aesthetics with Expert-Level Feedback Wan: Open and Advanced Large-Scale Video Generative Models

Reference 3

Resolution
metadata mismatch
local_arxiv, observed 2026-05-12T10:31:30.655756Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-02T06:30:47.504484+00:00.

source=pdf_text observed=2026-05-07T05:42:03.688350Z digest=sha256:4c1849651679e729c2419cd8dba8cbc3527f996b5ccd3848e1a3eeff7ae17b19

Observation 94c97aaf-bf80-4fdf-9ce0-9ebb3e7c264a · outbound

This paper cites DanceGRPO: Unleashing GRPO on Visual Generation.

AesRM: Improving Video Aesthetics with Expert-Level Feedback DanceGRPO: Unleashing GRPO on Visual Generation

Reference 4

Resolution
metadata mismatch
local_arxiv, observed 2026-05-12T10:36:28.613267Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-02T06:30:47.504484+00:00.

source=pdf_text observed=2026-05-07T05:42:03.688350Z digest=sha256:3183ab094b93e938d9b70ed85218367f36989a1836b5aba5b7d7f46645cb0478

Observation f1fc2eaf-e21a-4372-bd60-1af5bded01c1 · outbound

This paper cites InternVL3: Exploring Advanced Training and Test-Time Recipes for Open-Source Multimodal Models.

AesRM: Improving Video Aesthetics with Expert-Level Feedback InternVL3: Exploring Advanced Training and Test-Time Recipes for Open-Source Multimodal Models

Reference 5

Resolution
metadata mismatch
local_arxiv, observed 2026-05-12T10:31:30.661239Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-02T06:30:47.504484+00:00.

source=pdf_text observed=2026-05-07T05:42:03.688350Z digest=sha256:89cc0ca5475a6deda574dd7ade358779292f6b380aabb0454ede357fef3b9463

Observation efc9ea92-24a9-4507-995d-963ed4f52dfd · outbound

This paper cites an unresolved cited work.

AesRM: Improving Video Aesthetics with Expert-Level Feedback Unresolved cited work

Reference 6

Resolution
unresolved
raw_fallback, observed 2026-05-27T10:33:59.249091Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-02T06:30:47.504484+00:00.

source=pdf_text observed=2026-05-07T05:42:03.688350Z digest=sha256:cdc01d87e2c337d32168541e03a892d744d4a6d4581bb48012b31f03ec91da43

Observation abf03cca-e2bf-489a-9f7e-91e58043f210 · outbound

This paper cites an unresolved cited work.

AesRM: Improving Video Aesthetics with Expert-Level Feedback Unresolved cited work

Reference 7

Resolution
unresolved
raw_fallback, observed 2026-05-27T10:33:59.229010Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-02T06:30:47.504484+00:00.

source=pdf_text observed=2026-05-07T05:42:03.688350Z digest=sha256:053f25bb4ee5d75a78ac51282f57298989dd1bd9b3669070b4c571411a95ef84

Observation 77493d8e-81c4-40ba-8b13-404b0832129a · outbound

This paper cites an unresolved cited work.

AesRM: Improving Video Aesthetics with Expert-Level Feedback Unresolved cited work

Reference 8

Resolution
unresolved
raw_fallback, observed 2026-05-27T10:33:59.241399Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-02T06:30:47.504484+00:00.

source=pdf_text observed=2026-05-07T05:42:03.688350Z digest=sha256:e3f6ed55f6cd874f7463de49fe5e5546daee32bd9c56327eee4e8d91585903bf

Observation 29436fb8-e60f-4bf6-9663-120fb3dbbeee · outbound

This paper cites Light Style.

AesRM: Improving Video Aesthetics with Expert-Level Feedback Light Style

Reference 9

Resolution
malformed identifier
raw_fallback, observed 2026-05-27T10:33:59.237193Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-02T06:30:47.504484+00:00.

source=pdf_text observed=2026-05-07T05:42:03.688350Z digest=sha256:30c4af0eb1f7e172ade45aa2167f68d37e3eee851286231960852abb09956968

Observation 9c364800-1680-4876-b7b1-1ba02b646ea8 · outbound

This paper cites an unresolved cited work.

AesRM: Improving Video Aesthetics with Expert-Level Feedback Unresolved cited work

Reference 10

Resolution
unresolved
raw_fallback, observed 2026-05-27T10:33:59.233547Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-02T06:30:47.504484+00:00.

source=pdf_text observed=2026-05-07T05:42:03.688350Z digest=sha256:0308213f6d71ddfbf6fbbb0145865132fc8767a92e369bd2cc1e66110b3e510f

Observation 3f02a3bc-bd66-4760-9aae-05169ab3e327 · outbound

This paper cites For example: Video A underperforms Video B in visual aesthetics, while the two are comparable in visual fidelity and visual plausibility.

AesRM: Improving Video Aesthetics with Expert-Level Feedback For example: Video A underperforms Video B in visual aesthetics, while the two are comparable in visual fidelity and visual plausibility

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-05-27T10:33:59.245509Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-02T06:30:47.504484+00:00.

source=pdf_text observed=2026-05-07T05:42:03.688350Z digest=sha256:924c77bddf1c6c568758cddb730be2e9a26c10e7ad182a51aa76ca092446d084

Observation 13041c59-bbee-42de-8a5e-db9b6a804e64 · outbound

This paper cites Table 9: Full system prompt for AesRM-CoT.

AesRM: Improving Video Aesthetics with Expert-Level Feedback Table 9: Full system prompt for AesRM-CoT

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-05-27T10:33:59.225055Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-02T06:30:47.504484+00:00.

source=pdf_text observed=2026-05-07T05:42:03.688350Z digest=sha256:b69820acb2f0f4f223182ec1e7494f3239efa4c48130ddb443dad22287c11e3c

Observation 04d36ffe-70eb-44be-aaf7-5ecd1f68377a · outbound

This paper cites an unresolved cited work.

AesRM: Improving Video Aesthetics with Expert-Level Feedback Unresolved cited work

Reference 13

Resolution
unresolved
raw_fallback, observed 2026-05-27T10:33:59.255739Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-02T06:30:47.504484+00:00.

source=pdf_text observed=2026-05-07T05:42:03.688350Z digest=sha256:0225ad7d2e4617181c70b4b0fec8c252d70a8f7644c949a0def7a8be5207079e

Observation b5116620-74c8-4d58-8c27-726ae4c013ad · outbound

This paper cites an unresolved cited work.

AesRM: Improving Video Aesthetics with Expert-Level Feedback Unresolved cited work

Reference 14

Resolution
unresolved
raw_fallback, observed 2026-05-27T10:33:59.252750Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-02T06:30:47.504484+00:00.

source=pdf_text observed=2026-05-07T05:42:03.688350Z digest=sha256:4c846feaa6d06338362a21c57447e2841f4dc9aaceb3eab032d81f585ee12d52

Observation 7aa91d4d-b330-4471-be34-905b4a570640 · outbound

This paper cites $! & % #.

AesRM: Improving Video Aesthetics with Expert-Level Feedback $! & % #

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-05-27T10:33:59.221095Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-02T06:30:47.504484+00:00.

source=pdf_text observed=2026-05-07T05:42:03.688350Z digest=sha256:62563414637e8a78e52e4c6640a7e5fa971e0621cb02ea22b7c796136cbb3538

Pith citing papers

Observation 336ca822-3d25-4145-8f8e-12c4bc52fb9a · inbound

MSAVBench: Towards Comprehensive and Reliable Evaluation of Multi-Shot Audio-Video Generation cites this paper.

MSAVBench: Towards Comprehensive and Reliable Evaluation of Multi-Shot Audio-Video Generation AesRM: Improving Video Aesthetics with Expert-Level Feedback

Reference 21

Resolution
verified exact
local_arxiv, observed 2026-05-20T05:18:03.056737Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-02T06:30:47.504484+00:00.

source=pdf_text observed=2026-05-20T05:17:11.690484Z digest=sha256:d1c98b40e5bba46a7f69bb0bc4bc5cc4688029463a56f7311c246c7c3eb4185b

Observation 073be41a-85d1-4974-833e-5255aec4d323 · inbound

MSAVBench: Towards Comprehensive and Reliable Evaluation of Multi-Shot Audio-Video Generation cites this paper.

MSAVBench: Towards Comprehensive and Reliable Evaluation of Multi-Shot Audio-Video Generation AesRM: Improving Video Aesthetics with Expert-Level Feedback

Reference 21

Resolution
verified exact
local_arxiv, observed 2026-06-30T18:04:58.055663Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-02T06:30:47.504484+00:00.

source=pdf_text observed=2026-06-30T18:00:39.556424Z digest=sha256:7e3d199800d95c51ac102db6954c8f7cafe20f6ca6543ab46b537ea59c918188