Pith. sign in

Paper Citation Record · LEDGER

From Sound to Sight: Towards AI-authored Music Videos

As of 11 August 2026, this Paper Citation Record lists 67 of 67 outbound references and 0 inbound Pith citation observations for arXiv:2509.00029.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2509.00029 v1

Coverage vector

measured 67 of 67 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-05T18:23:31.599121Z

measured 67 of 67 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-11T06:34:44.6726+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

67 of 67 outbound references displayed

  • verified exact3
  • verified fuzzy49
  • unresolved15
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 35b71881-cbfc-4364-bd1e-bc64100746e9 · outbound

This paper cites Secure & Personalized Music-to- Video Generation via CHARCHA, 2025.

From Sound to Sight: Towards AI-authored Music Videos Secure & Personalized Music-to- Video Generation via CHARCHA, 2025

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T18:23:42.054602Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-05T18:23:26.897893Z digest=sha256:2700325a9f8fe10357a63de12e5a977575795f7daedc9ac29d57177501acfb17

Observation a61a3c3b-bbed-48d2-94ba-e2d358488d00 · outbound

This paper cites ImproveYourVideos: Architectural Im- provements for Text-to-Video Generation Pipeline.IEEE Ac- cess, 13:1986–2003, 2025.

From Sound to Sight: Towards AI-authored Music Videos ImproveYourVideos: Architectural Im- provements for Text-to-Video Generation Pipeline.IEEE Ac- cess, 13:1986–2003, 2025

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T18:23:41.898128Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-05T18:23:26.959917Z digest=sha256:b43ab447096a8d3811003177aa2d53dda3e6f7ed92ea03cc89d4bdbb3f7fd011

Observation e8302f5f-55e2-4c94-90e2-e384f004d410 · outbound

This paper cites The MIT Press, Cambridge, Massachusetts, 2021.

From Sound to Sight: Towards AI-authored Music Videos The MIT Press, Cambridge, Massachusetts, 2021

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T18:23:41.702044Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-05T18:23:27.037973Z digest=sha256:8e3c544d486d2067f99840184ee500a9a0fddcb2a58954782ad9b848230b7dc0

Observation 5ba3e4db-b1de-4f28-aca2-68fa68d69f68 · outbound

This paper cites Boden and Ernest A.

From Sound to Sight: Towards AI-authored Music Videos Boden and Ernest A

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T18:23:41.545856Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-05T18:23:27.112468Z digest=sha256:78b73a723d8c88b23f79d41c411fa27a0f29c4276bd6573e9009086f4cd0dd4a

Observation bdad2ee3-a30d-475a-b094-ded1e023e0cd · outbound

This paper cites Review of Gottschall (2012): The story- telling animal: How stories make us human.Scientific Study of Literature, 2(2):317–321, 2012.

From Sound to Sight: Towards AI-authored Music Videos Review of Gottschall (2012): The story- telling animal: How stories make us human.Scientific Study of Literature, 2(2):317–321, 2012

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T18:23:41.416285Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-05T18:23:27.193693Z digest=sha256:69a69828f7ce26a3e189b302dfcb483c93f10cab78e571fd028161fc8e972025

Observation c291cf6f-437e-4aab-912a-8393c68895a1 · outbound

This paper cites Diffusion Models as Artists: Are we Closing the Gap between Humans and Machines?.

From Sound to Sight: Towards AI-authored Music Videos Diffusion Models as Artists: Are we Closing the Gap between Humans and Machines?

Reference 6

Resolution
verified exact
local_arxiv, observed 2026-08-05T18:23:32.266214Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-05T18:23:27.331853Z digest=sha256:7ef654ea2094dddf0f5f9c8ce98b81f19c75b10b1e87f485892c5d43f8a08efc

Observation 4c861e84-251a-4ec6-a422-bb2a7f26388b · outbound

This paper cites Crossmodal associations between naturally occurring tactile and sound textures.Per- ception, 53(4):219–239, 2024.

From Sound to Sight: Towards AI-authored Music Videos Crossmodal associations between naturally occurring tactile and sound textures.Per- ception, 53(4):219–239, 2024

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T18:23:41.268228Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-05T18:23:27.415389Z digest=sha256:ebe24b3d89b10c434a4c71e0aa914e6d2412c0825acf9225cf380a3373c7cd02

Observation 1383f726-cf8c-4abb-b8f7-54389b6414c0 · outbound

This paper cites Cancino-Chac ´on, Maarten Grachten, Werner Goebl, and Gerhard Widmer.

From Sound to Sight: Towards AI-authored Music Videos Cancino-Chac ´on, Maarten Grachten, Werner Goebl, and Gerhard Widmer

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T18:23:41.039657Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-05T18:23:27.487946Z digest=sha256:33050bec9ff3df3750f048ddfe8abf794cb54ba56ebdf3de60bf242082fb295d

Observation edca7429-49e1-4a90-b226-69ee5cb045ed · outbound

This paper cites ”scary robots”: Examining public responses to ai.

From Sound to Sight: Towards AI-authored Music Videos ”scary robots”: Examining public responses to ai

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T18:23:40.868573Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-05T18:23:27.566047Z digest=sha256:c12930aa2425c749a63dec1866d04f0e1bc6826c0c1a0869e307462ba43bc821

Observation b0ca1abb-0212-4b0c-b9df-58d12a4c64f9 · outbound

This paper cites Understanding and Creating Art with AI: Review and Outlook.

From Sound to Sight: Towards AI-authored Music Videos Understanding and Creating Art with AI: Review and Outlook

Reference 10

Resolution
verified exact
local_arxiv, observed 2026-08-05T18:23:32.063832Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-05T18:23:27.698936Z digest=sha256:273767e7e7981e9278769310993d3a2f57c693be4f4b2e9a29a5f6476cc47874

Observation cebe88bd-f23b-4be7-8dd7-154a2ad1f1e6 · outbound

This paper cites Dasovich-Wilson, Marc Thompson, and Suvi Saarikallio.

From Sound to Sight: Towards AI-authored Music Videos Dasovich-Wilson, Marc Thompson, and Suvi Saarikallio

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T18:23:40.679610Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-05T18:23:27.784957Z digest=sha256:ac7c91b43ddf7da8ec7aabddb04a0fd942d7e10d507324125be6e64293f55ec8

Observation 54998d99-14d2-4bad-bb6b-7cf9c35eff00 · outbound

This paper cites DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning.

From Sound to Sight: Towards AI-authored Music Videos DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-05T18:23:27.845221Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T18:23:27.845221Z digest=sha256:7edb727011960000565c3066cb1ee5f5df32e49c7732ff8d119af30875565303

Observation 2b292f5d-f0b2-440a-b3b8-14ef5615f0c2 · outbound

This paper cites Pengi: an audio language model for audio tasks.

From Sound to Sight: Towards AI-authored Music Videos Pengi: an audio language model for audio tasks

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T18:23:40.508262Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-05T18:23:27.892730Z digest=sha256:ddc4775550f49d66210bb410c4bd015c0e0d23437626038d236beca3ed15d79e

Observation 5438db99-ddb5-4b7e-b188-6f788828f549 · outbound

This paper cites Berkeley Publishing Group, New York, New York, 2005.

From Sound to Sight: Towards AI-authored Music Videos Berkeley Publishing Group, New York, New York, 2005

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T18:23:40.316719Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-05T18:23:27.988577Z digest=sha256:15255cf8c86ac4d9b93837a7514450095d2c5337056710f3fe5ece906a55edc0

Observation ec55e3db-b401-425e-a0c1-e5f396de7d9e · outbound

This paper cites EasyVid AI Video Maker, 2024.

From Sound to Sight: Towards AI-authored Music Videos EasyVid AI Video Maker, 2024

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T18:23:40.146291Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-05T18:23:28.068936Z digest=sha256:b7c571c71887df338bfdf7a748c6099f43fea5f3608e4542998c477920322067

Observation f7ef522e-3c84-4be7-90a0-ec5035bf86e8 · outbound

This paper cites CAN: Creative Adversarial Networks, Generating "Art" by Learning About Styles and Deviating from Style Norms.

From Sound to Sight: Towards AI-authored Music Videos CAN: Creative Adversarial Networks, Generating "Art" by Learning About Styles and Deviating from Style Norms

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-05T18:23:28.167494Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T18:23:28.167494Z digest=sha256:6e34a617ef8e72f3432fa5ee0c80848adde2859329ec02ad92658085b1a1f978

Observation 690b92dd-d1ba-415f-b0e0-47f770d65da6 · outbound

This paper cites CLAP Learning Audio Concepts from Natural Language Supervision.ICASSP 2023 - 2023 IEEE International Conference on Acoustics, Speech and Signal Processing (ICASSP), pages 1–5, 2023.

From Sound to Sight: Towards AI-authored Music Videos CLAP Learning Audio Concepts from Natural Language Supervision.ICASSP 2023 - 2023 IEEE International Conference on Acoustics, Speech and Signal Processing (ICASSP), pages 1–5, 2023

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T18:23:39.930859Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-05T18:23:28.237066Z digest=sha256:8d174f7ca090a474c543026dde37e8118bcf010c806610c6c942bef91e20453e

Observation d0d87f1c-8fa1-40e8-9acc-a015489e44bd · outbound

This paper cites Rand, and Iyad Rah- wan.

From Sound to Sight: Towards AI-authored Music Videos Rand, and Iyad Rah- wan

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T18:23:39.659126Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-05T18:23:28.330599Z digest=sha256:7b81595412dcc30c26eb0685f43a10c8f32061380058328a7de102860859a926

Observation 28a48ced-bc6f-4edc-a4ae-bd6ef9e7845e · outbound

This paper cites Frank, Matthew Groh, Laura Herman, Neil Leach, Robert Mahari, Alex “Sandy” Pentland, Olga Russakovsky, Hope Schroeder, and Amy Smith.

From Sound to Sight: Towards AI-authored Music Videos Frank, Matthew Groh, Laura Herman, Neil Leach, Robert Mahari, Alex “Sandy” Pentland, Olga Russakovsky, Hope Schroeder, and Amy Smith

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T18:23:39.392080Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-05T18:23:28.455570Z digest=sha256:811df81a2d50da12d8d51b7583834f339def257f035e0c2b8cc5a00804607112

Observation b03fa92e-d669-4b4c-b3cc-618d47076bdd · outbound

This paper cites GAMA: A Large Audio-Language Model with Advanced Audio Understanding and Complex Reasoning Abilities.

From Sound to Sight: Towards AI-authored Music Videos GAMA: A Large Audio-Language Model with Advanced Audio Understanding and Complex Reasoning Abilities

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-05T18:23:28.522442Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T18:23:28.522442Z digest=sha256:f79668e3ce1b39af49015366d2cf1c593fccbbe714cfdc01f9b03a44b844337c

Observation 6a5f500d-6737-4254-a0b2-d12bd44783da · outbound

This paper cites From Ragtime to Swingtime: Fifty Glittering Years of Stage and Song.

From Sound to Sight: Towards AI-authored Music Videos From Ragtime to Swingtime: Fifty Glittering Years of Stage and Song

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T18:23:39.058349Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-05T18:23:28.614731Z digest=sha256:5a4ac33f3e1542a1d1903f67ab263a2f62d25ecaed3adb81e7e125b7247c4738

Observation cb10c913-26cb-4406-ba05-edefb477c907 · outbound

This paper cites Generative adversarial networks.Commu- nications of the ACM, 63(11):139–144, 2020.

From Sound to Sight: Towards AI-authored Music Videos Generative adversarial networks.Commu- nications of the ACM, 63(11):139–144, 2020

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-05T18:23:28.676982Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T18:23:28.676982Z digest=sha256:25bbe39e96c145ce05423a820abb346eb7d64d12e2493dcd2a4c07d500562502

Observation d5233c9e-2fff-4183-9f49-a040759ab52c · outbound

This paper cites Graesser, Murray Singer, and Tom Trabasso.

From Sound to Sight: Towards AI-authored Music Videos Graesser, Murray Singer, and Tom Trabasso

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T18:23:38.781115Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-05T18:23:28.760858Z digest=sha256:92f907817509b3fddf4635ae01fdbf31fc8258dc712f0b208e11c77c460db015

Observation 8e3b659f-8a81-491c-9cbb-9245ba06337e · outbound

This paper cites Beware of fictional ai narratives.Nature Machine Intelligence, 2(11):654–654, 2020.

From Sound to Sight: Towards AI-authored Music Videos Beware of fictional ai narratives.Nature Machine Intelligence, 2(11):654–654, 2020

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T18:23:38.617464Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-05T18:23:28.852552Z digest=sha256:0c91e62345065042d5c19425350fc8792356e1c9955a3b2d5487b4675956b97c

Observation c7a6a6e8-13fa-4429-99e4-29a1ce06a8b9 · outbound

This paper cites Can Computers Create Art?.

From Sound to Sight: Towards AI-authored Music Videos Can Computers Create Art?

Reference 25

Resolution
verified exact
local_arxiv, observed 2026-08-05T18:23:31.876234Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-05T18:23:28.947855Z digest=sha256:db71170dce912a7c08552b1b787b817977410766a8bc28d4ad730da6dd3ad83b

Observation 88c7f4d6-e73e-4761-89ed-973b8d6cd02f · outbound

This paper cites Computers do not make art, people do.

From Sound to Sight: Towards AI-authored Music Videos Computers do not make art, people do

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T18:23:38.444000Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-05T18:23:29.024474Z digest=sha256:df90478fa8eb9b38ff831936e055a6838abcf97b0d07c9ce200a5c1faa9dc4c6

Observation 8474ac28-c3e5-4463-9bb9-edf5a15669f1 · outbound

This paper cites Cascaded Diffusion Models for High Fidelity Image Generation.

From Sound to Sight: Towards AI-authored Music Videos Cascaded Diffusion Models for High Fidelity Image Generation

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-05T18:23:29.137643Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T18:23:29.137643Z digest=sha256:5304a35f1986df68a939683a6351cfa7c8e1416195399b77d021cb781a0d1966

Observation 2ece083d-74c1-4451-ab72-1d3abc6687e3 · outbound

This paper cites Video Diffusion Models.

From Sound to Sight: Towards AI-authored Music Videos Video Diffusion Models

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-05T18:23:29.193146Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T18:23:29.193146Z digest=sha256:db467c60b12a7d6b00ab135648779e0c006f9c33f54d79deb75f4df37c47e9ce

Observation 4240c902-f576-46d6-a20b-3c28a669161d · outbound

This paper cites Artificial Intelli- gence, Artists, and Art: Attitudes Toward Artwork Produced by Humans vs.

From Sound to Sight: Towards AI-authored Music Videos Artificial Intelli- gence, Artists, and Art: Attitudes Toward Artwork Produced by Humans vs

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T18:23:38.287509Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-05T18:23:29.276918Z digest=sha256:040c12b7957a3aa660a3e622af2b474f9498bd85a1544b15889b242d59408f25

Observation 503c4c33-4102-42e1-8021-e4b0c846cba7 · outbound

This paper cites Blaine Horton Jr, Michael W.

From Sound to Sight: Towards AI-authored Music Videos Blaine Horton Jr, Michael W

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T18:23:38.054367Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-05T18:23:29.328957Z digest=sha256:93b8eca08fc18ff886e39bf98144a6016378b9664f1a408028fd90568ac60fe4

Observation 41388663-6713-471c-ac63-3afa9d8e79ef · outbound

This paper cites VBench: Com- prehensive Benchmark Suite for Video Generative Models,.

From Sound to Sight: Towards AI-authored Music Videos VBench: Com- prehensive Benchmark Suite for Video Generative Models,

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-05T18:23:29.398758Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T18:23:29.398758Z digest=sha256:7caf736260720f65413ad542fef30250ed3b93456091b3c0b572a1796298b597

Observation 6b1ec288-881f-4ad2-baf4-f984b31c6e91 · outbound

This paper cites Cross-modal associations between paintings and sounds: Ef- fects of embodiment.Perception, 51(12):871–888, 2022.

From Sound to Sight: Towards AI-authored Music Videos Cross-modal associations between paintings and sounds: Ef- fects of embodiment.Perception, 51(12):871–888, 2022

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T18:23:37.848811Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-05T18:23:29.573162Z digest=sha256:a024fa40df3e7bbc8d32af3590f6eba39cd08a6b3607fc36666825083b56a3c6

Observation 15bb7bbc-1eac-47d0-83b2-3fd9b6249836 · outbound

This paper cites Kaiber AI: Generating Videos with Superstu- dio, 2025.

From Sound to Sight: Towards AI-authored Music Videos Kaiber AI: Generating Videos with Superstu- dio, 2025

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T18:23:37.644475Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-05T18:23:29.639115Z digest=sha256:4b10e87b36f7b889e832e1f6891b847e17b452498dc5528b58aef219eeea9cf9

Observation 8704e49a-128b-44ba-a303-d5f9636c681a · outbound

This paper cites Artificial Intelligence and Copyright: Le- gal Quandary in the Digital Age: Some Musings.SSRN Elec- tronic Journal, 2021.

From Sound to Sight: Towards AI-authored Music Videos Artificial Intelligence and Copyright: Le- gal Quandary in the Digital Age: Some Musings.SSRN Elec- tronic Journal, 2021

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T18:23:37.465831Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-05T18:23:29.678001Z digest=sha256:0559373eb859f3f2741bb841b1f78a3e019548c37b0252e89716d05f08c675a9

Observation 698aae6f-f86f-4f8e-b48d-636c01161c93 · outbound

This paper cites ”AI enhances our performance, I have no doubt this one will do the same”: The Placebo effect is ro- bust to negative descriptions of AI.

From Sound to Sight: Towards AI-authored Music Videos ”AI enhances our performance, I have no doubt this one will do the same”: The Placebo effect is ro- bust to negative descriptions of AI

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T18:23:37.288366Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-05T18:23:29.734159Z digest=sha256:ef3589beb8a5955314ed7b71ec945826a2e0e768ef03a0a18108236b93e6f1d0

Observation 5fa8b5dc-de0b-45a2-b11c-af5887acc0f9 · outbound

This paper cites The (R)evolution of Music Video in American Music Industry.New Horizons in English Stud- ies, 8:163–176, 2023.

From Sound to Sight: Towards AI-authored Music Videos The (R)evolution of Music Video in American Music Industry.New Horizons in English Stud- ies, 8:163–176, 2023

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T18:23:37.075736Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-05T18:23:29.780560Z digest=sha256:012d0aed9706bed24881f2f4d6a334be9145c43f25dfc6a1e9a809f909d138c5

Observation 022a4c7f-7d85-4e6e-9fd6-3b600fcd24f6 · outbound

This paper cites The Placebo Effect of Artificial Intelli- gence in Human–Computer Interaction.ACM Transactions on Computer-Human Interaction, 29(6):1–32, 2022.

From Sound to Sight: Towards AI-authored Music Videos The Placebo Effect of Artificial Intelli- gence in Human–Computer Interaction.ACM Transactions on Computer-Human Interaction, 29(6):1–32, 2022

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T18:23:36.853711Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-05T18:23:29.812172Z digest=sha256:2574e04d6688d9069ce7e7fae08e7fe94758556931e626b819bc7b018a69cb8c

Observation 4be64de3-089b-4070-8776-a39526b934c4 · outbound

This paper cites Robust One Shot Audio to Video Generation.

From Sound to Sight: Towards AI-authored Music Videos Robust One Shot Audio to Video Generation

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T18:23:36.615800Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-05T18:23:29.912793Z digest=sha256:e9b6beb79b5d0f572123492da2cf22ba4643f681b2ddefb381973af1bfc4fbf7

Observation 74bd9a09-60a2-4fcf-95e4-2f5fcce6270f · outbound

This paper cites an unresolved cited work.

From Sound to Sight: Towards AI-authored Music Videos Unresolved cited work

Reference 39

Resolution
unresolved
raw_fallback, observed 2026-08-05T18:23:36.408063Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-05T18:23:29.975981Z digest=sha256:da5bc2d7a857ad4e1cda6a0c5d1012a78c1d9845c895ba1f805d82f2c44f934b

Observation 3d797bd9-75fa-468f-a6c2-3c4459d4bd54 · outbound

This paper cites Lima, Carlos G.

From Sound to Sight: Towards AI-authored Music Videos Lima, Carlos G

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T18:23:36.189412Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-05T18:23:30.043470Z digest=sha256:5a3dea087e667d3e6ebfde7d714e136cd74202abcbcf6e0c1ef751f027dc8bfc

Observation ac8a5eb0-6a88-4dae-92d6-665e84ccc9e7 · outbound

This paper cites In ai we trust? effects of agency locus and transparency on uncertainty reduction in human–ai interac- tion.Journal of Computer-Mediated Communication, 26(6): 384–402, 2021.

From Sound to Sight: Towards AI-authored Music Videos In ai we trust? effects of agency locus and transparency on uncertainty reduction in human–ai interac- tion.Journal of Computer-Mediated Communication, 26(6): 384–402, 2021

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T18:23:35.961495Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-05T18:23:30.099710Z digest=sha256:92337aa709c7869672c43f3bab5efe0074bb3e961ec6211be6f6cb7dad6a2661

Observation 3230a59a-0e88-4646-b8e0-6bce11a5b60c · outbound

This paper cites Revisiting the Gold Standard: Grounding Summarization Evaluation with Robust Human Evaluation.

From Sound to Sight: Towards AI-authored Music Videos Revisiting the Gold Standard: Grounding Summarization Evaluation with Robust Human Evaluation

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T18:23:35.771559Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-05T18:23:30.149947Z digest=sha256:d56c1d79cb3916627b602b4f78c4f467c7adff879797b16ce822a5a9ba48c945

Observation f95758a0-0e62-46ba-b3df-1982fbfe6de0 · outbound

This paper cites Are Emergent Abilities in Large Language Models just In-Context Learning?.

From Sound to Sight: Towards AI-authored Music Videos Are Emergent Abilities in Large Language Models just In-Context Learning?

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-05T18:23:30.219156Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T18:23:30.219156Z digest=sha256:bcadf04d7c0f902f19ab1454d51d4009d8a5e728bd20bc904b6cc3b132da6805

Observation 00238615-afab-462f-8c81-689dc8f2f935 · outbound

This paper cites Art, Creativity, and the Potential of Artificial Intelligence.Arts, 8(1):26, 2019.

From Sound to Sight: Towards AI-authored Music Videos Art, Creativity, and the Potential of Artificial Intelligence.Arts, 8(1):26, 2019

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T18:23:35.597409Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-05T18:23:30.292545Z digest=sha256:1e3d7f2b9b4a6c25909abfdc09bae568e3d8c465fb9e9c6b511d944f04f8805a

Observation 538e82b9-738e-4623-80bd-8ac68965494d · outbound

This paper cites Windows Media Player, 2025.

From Sound to Sight: Towards AI-authored Music Videos Windows Media Player, 2025

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T18:23:35.424087Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-05T18:23:30.365220Z digest=sha256:bf74730c445d441776d12042d3ecdae2dfb2fd6578dd1f09f1cf0b1b2cad353a

Observation d9d86d38-155b-4fab-9b62-ab2753341979 · outbound

This paper cites Com- putational Music Structure Analysis (Dagstuhl Seminar 16092).

From Sound to Sight: Towards AI-authored Music Videos Com- putational Music Structure Analysis (Dagstuhl Seminar 16092)

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T18:23:35.253847Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-05T18:23:30.413253Z digest=sha256:a588e8c451f8051a2d23d29d25515194d41388c763f06279e644ce1c5dc32f69

Observation 8439a621-4d74-40e0-b9f8-c498a573d51c · outbound

This paper cites AI Music Video Generator, 2025.

From Sound to Sight: Towards AI-authored Music Videos AI Music Video Generator, 2025

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T18:23:35.081445Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-05T18:23:30.475329Z digest=sha256:c17287deb199a4f41984b29c3a865a065a19c73efd39039ee72e0c65c710ee41

Observation 4b2417f8-7576-4bb6-a7bd-c4c2be20e346 · outbound

This paper cites Hello GPT-4o by OpenAI, 2025.

From Sound to Sight: Towards AI-authored Music Videos Hello GPT-4o by OpenAI, 2025

Reference 48

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T18:23:34.888134Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-05T18:23:30.514243Z digest=sha256:3706e323a6d6bfbf2a353a4fdf0ebbcf21489b465e72915b2e724a12cfc760bf

Observation d06b6828-bac5-48b9-92cf-986da4e8dcf0 · outbound

This paper cites Synaesthesia.European Neurology, 57(2):120– 124, 2007.

From Sound to Sight: Towards AI-authored Music Videos Synaesthesia.European Neurology, 57(2):120– 124, 2007

Reference 49

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T18:23:34.657679Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-05T18:23:30.606125Z digest=sha256:c94671d0c7fa7a7714d4b757039ecc1d59bdd5b6504e9bf8225a15ddf5bfe599

Observation 4476f215-5682-4707-a602-e679e3427c66 · outbound

This paper cites Robust Speech Recognition via Large-Scale Weak Supervision,.

From Sound to Sight: Towards AI-authored Music Videos Robust Speech Recognition via Large-Scale Weak Supervision,

Reference 50

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T18:23:34.467417Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-05T18:23:30.667108Z digest=sha256:4134c84a3109598ebda7d47377dd34e4e25e7639594b2e3cad2d8f20e08133b8

Observation 1e2aeb90-a502-4ce8-9300-b14926c3376d · outbound

This paper cites Self-supervised Dance Video Synthesis Conditioned on Mu- sic.

From Sound to Sight: Towards AI-authored Music Videos Self-supervised Dance Video Synthesis Conditioned on Mu- sic

Reference 51

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T18:23:34.356357Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-05T18:23:30.785288Z digest=sha256:2891727aaf7ca0e1c7f6011a0add9e9623a2828a44787bbafb43f0f950e0864d

Observation 932469f3-9384-422f-b1db-2cbf574f3cf6 · outbound

This paper cites High-Resolution Image Synthesis With Latent Diffusion Models.

From Sound to Sight: Towards AI-authored Music Videos High-Resolution Image Synthesis With Latent Diffusion Models

Reference 52

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T18:23:34.209998Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-05T18:23:30.911969Z digest=sha256:e502752d94b64687b4a1599a8bc3cc3f61c98f9e57bbe20525f137440b285ef6

Observation c9259e24-855a-4884-84ad-ea6c0f4f8482 · outbound

This paper cites Oxford University PressNew York, NY , 1995.

From Sound to Sight: Towards AI-authored Music Videos Oxford University PressNew York, NY , 1995

Reference 53

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T18:23:34.033741Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-05T18:23:30.966040Z digest=sha256:171aad38114b33eec93459f77ec7087c2c7ab531a599c1960379ad4a44a8a617

Observation 329a479d-5675-4df6-b4d0-b612c99fed59 · outbound

This paper cites Machine Learning Processes As Sources of Ambiguity: Insights from AI Art.

From Sound to Sight: Towards AI-authored Music Videos Machine Learning Processes As Sources of Ambiguity: Insights from AI Art

Reference 54

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T18:23:33.859082Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-05T18:23:31.056150Z digest=sha256:7eb34a2d3e7da84d76ccf15410d6e6afee498151761cd8c9421e836c09e198d3

Observation 26ef0629-2ced-490d-b627-9b126447c808 · outbound

This paper cites Mochi 1 by Genmo Team, 2024.

From Sound to Sight: Towards AI-authored Music Videos Mochi 1 by Genmo Team, 2024

Reference 55

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T18:23:33.748561Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-05T18:23:31.097961Z digest=sha256:b7b00d1870e092336b2d8a440c959cdbbe21070a5b9f5cb007a0e32b945be41b

Observation 6efd3b9f-d5b5-4315-8054-5d2f427bca2f · outbound

This paper cites Revid.ai, 2025.

From Sound to Sight: Towards AI-authored Music Videos Revid.ai, 2025

Reference 56

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T18:23:33.639793Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-05T18:23:31.151817Z digest=sha256:c56a1a8a88686ea9d0b983ea426bc2ad894bd8365af881d667d14608ebe597f0

Observation 14680611-1f71-4210-b9cf-19d4446e1bcf · outbound

This paper cites Specterr: Music Video Maker Online, 2025.

From Sound to Sight: Towards AI-authored Music Videos Specterr: Music Video Maker Online, 2025

Reference 57

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T18:23:33.425690Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-05T18:23:31.179963Z digest=sha256:8720df46bdbefd974c76f1e2da508d2f244402767bcf54d661aff87c34007a59

Observation e2ae8bdf-dd3f-46a0-813c-da86bef81a7f · outbound

This paper cites Wan: Open and Advanced Large-Scale Video Generative Models.

From Sound to Sight: Towards AI-authored Music Videos Wan: Open and Advanced Large-Scale Video Generative Models

Reference 58

Resolution
unresolved
no resolver link, observed 2026-08-05T18:23:31.222822Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T18:23:31.222822Z digest=sha256:e22cf09506d4430892295a26a236627397832794d4ea364644414b77eff52728

Observation 41e18c0c-b122-424c-b254-b67062be121a · outbound

This paper cites Chain-of-Thought Prompting Elicits Reasoning in Large Language Models.

From Sound to Sight: Towards AI-authored Music Videos Chain-of-Thought Prompting Elicits Reasoning in Large Language Models

Reference 59

Resolution
unresolved
no resolver link, observed 2026-08-05T18:23:31.278343Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T18:23:31.278343Z digest=sha256:e8db84af0e52d3cb26d06e78e0ae2731112060889c0b45ab40088995eed0175a

Observation e7073e7c-785f-4103-891a-2cb15cfe1d07 · outbound

This paper cites A Survey on Knowledge Distillation of Large Language Models.

From Sound to Sight: Towards AI-authored Music Videos A Survey on Knowledge Distillation of Large Language Models

Reference 60

Resolution
unresolved
no resolver link, observed 2026-08-05T18:23:31.358582Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T18:23:31.358582Z digest=sha256:525d1311931a8cbf6f94618a1e842beaa40c0fe25c42d75aba52d3752e3d33ab

Observation 49183d08-19ef-4db6-8e03-37ab707352bf · outbound

This paper cites Wordcraft: Story Writing With Large Language Models.

From Sound to Sight: Towards AI-authored Music Videos Wordcraft: Story Writing With Large Language Models

Reference 61

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T18:23:33.201784Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-05T18:23:31.419333Z digest=sha256:8354bc816878085776b4ee5c8122e08e6ea7044299e015a73bf6809835a3526f

Observation 9634ef36-91ee-42ab-a47c-4f8eb2540052 · outbound

This paper cites Let’s Play Music: Audio-Driven Performance Video Generation.

From Sound to Sight: Towards AI-authored Music Videos Let’s Play Music: Audio-Driven Performance Video Generation

Reference 62

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T18:23:33.006421Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-05T18:23:31.483700Z digest=sha256:37a6cc10ed8597907a8867236e20d389acf0115e9390c5c002adbf98cec6cbd6

Observation 298bad70-e43a-43d4-9bcb-25f33bafe350 · outbound

This paper cites SCENE #:.

From Sound to Sight: Towards AI-authored Music Videos SCENE #:

Reference 66

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T18:23:32.741767Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-05T18:23:31.551008Z digest=sha256:575f1a4bc1c651c1c1d2fabe9596a16277ffc9bec2985b447f55013f77919794

Observation 49386794-f5d9-4d91-b458-d8ee3e6eee88 · outbound

This paper cites All items were rated on a 7-point Likert scale: where 1 = Strongly Disagree, and 7 = Strongly Agree.

From Sound to Sight: Towards AI-authored Music Videos All items were rated on a 7-point Likert scale: where 1 = Strongly Disagree, and 7 = Strongly Agree

Reference 67

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T18:23:32.510045Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-05T18:23:31.599121Z digest=sha256:67b73ddbdd0dc26b06de4ef179317829bb9364566a59d72fca21dee0435f561e

Observation 93003541-2ae2-4dd4-8393-3136a5382996 · outbound

This paper cites an unresolved cited work.

From Sound to Sight: Towards AI-authored Music Videos Unresolved cited work

Reference 2020

Resolution
unresolved
no resolver link, observed 2026-08-05T18:23:30.852381Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T18:23:30.852381Z digest=sha256:dc0dc7a00563730f8c01c43637165d01a09993c5ca3850aa2a7d6047fda1063e

Observation c7a423f1-b732-419d-a460-ae3e23fc4ffa · outbound

This paper cites Robust Speech Recognition via Large-Scale Weak Supervision.

From Sound to Sight: Towards AI-authored Music Videos Robust Speech Recognition via Large-Scale Weak Supervision

Reference 2022

Resolution
unresolved
no resolver link, observed 2026-08-05T18:23:30.736114Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T18:23:30.736114Z digest=sha256:47ff8dc127806af917faf92cb10d0250bf95fb38e26ea16d45b24ffc2b8d83d0

Observation 24b92ba4-8eb7-4a3e-897e-9e3f3a89ca43 · outbound

This paper cites VBench: Comprehensive Benchmark Suite for Video Generative Models.

From Sound to Sight: Towards AI-authored Music Videos VBench: Comprehensive Benchmark Suite for Video Generative Models

Reference 2023

Resolution
unresolved
no resolver link, observed 2026-08-05T18:23:29.510645Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T18:23:29.510645Z digest=sha256:2479fb96903542d29882b8a0df33847abe268e1d86d02d2f7f21d72174e1113c

Pith citing papers

No inbound Pith citation observations are available.