Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-10T00:02:26.319596Z
Paper Citation Record · LEDGER
As of 13 August 2026, this Paper Citation Record lists 68 of 68 outbound references and 2 inbound Pith citation observations for arXiv:2501.18314.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-10T00:02:26.319596Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-13T06:32:02.005865+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-06T15:29:05.574108Z
A source-named dated measurement, never combined with another source.
Source: pith, observed 2026-08-06T05:00:19.327534Z
68 of 68 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation fdf31bf0-5965-4ee2-8d37-07a1a9098af5 · outbound
AGAV-Rater: Adapting Large Multimodal Model for AI-Generated Audio-Visual Quality Assessment Prime Voice AI, 2023
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 7c0b15e2-e613-4299-bc34-ce60acd262d7 · outbound
AGAV-Rater: Adapting Large Multimodal Model for AI-Generated Audio-Visual Quality Assessment Accessed December 28, 2023 [Online] https://www.pika.art/, 2023
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 21cbe542-4314-4966-a68f-7f2eb23af791 · outbound
AGAV-Rater: Adapting Large Multimodal Model for AI-Generated Audio-Visual Quality Assessment Accessed June 17, 2024 [Online] https://runwayml.com/research/introducing-gen-3-alpha, 2024
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 46e2f2ee-4c14-4881-a339-80f2d8423dca · outbound
AGAV-Rater: Adapting Large Multimodal Model for AI-Generated Audio-Visual Quality Assessment Accessed June 6, 2024 [Online] https://klingai.kuaishou.com/, 2024
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 92a46a10-49f5-4508-8728-34d5a329ffb4 · outbound
AGAV-Rater: Adapting Large Multimodal Model for AI-Generated Audio-Visual Quality Assessment Accessed February 16, 2024 [Online] https://openai.com/sora/, 2024
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 67604226-62bc-4995-8c68-4cb83de72b8f · outbound
AGAV-Rater: Adapting Large Multimodal Model for AI-Generated Audio-Visual Quality Assessment MusicLM: Generating Music From Text
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 54079295-2452-42d4-aea7-2fe544a8b568 · outbound
AGAV-Rater: Adapting Large Multimodal Model for AI-Generated Audio-Visual Quality Assessment G., Schmidmer, C., Berger, J., Obermann, M., Ullmann, R., Pomy, J., and Keyhl, M
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation b5aba429-cc3e-46f8-adc6-3f7f1351b3d9 · outbound
AGAV-Rater: Adapting Large Multimodal Model for AI-Generated Audio-Visual Quality Assessment Attention-guided neural networks for full-reference and no-reference audio-visual quality assessment
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 0b5b111c-0ede-4c07-8e30-3bce1046ada1 · outbound
AGAV-Rater: Adapting Large Multimodal Model for AI-Generated Audio-Visual Quality Assessment Subjective and objective audio-visual quality assessment for user generated content
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation f8009c33-1708-4b14-8d49-2048beb884b1 · outbound
AGAV-Rater: Adapting Large Multimodal Model for AI-Generated Audio-Visual Quality Assessment Semantically consistent Video-to-Audio Generation using Multimodal Language Large Model
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9f910086-50e1-41d5-bf02-9b05428547b6 · outbound
AGAV-Rater: Adapting Large Multimodal Model for AI-Generated Audio-Visual Quality Assessment Vggsound: A large-scale audio-visual dataset
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation f0af9eec-e1f8-4f70-a573-8304309ffba5 · outbound
AGAV-Rater: Adapting Large Multimodal Model for AI-Generated Audio-Visual Quality Assessment Vast: A vision-audio-subtitle-text omni-modality foundation model and dataset
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 9da84270-4d0e-43ae-a854-9f56b0425268 · outbound
AGAV-Rater: Adapting Large Multimodal Model for AI-Generated Audio-Visual Quality Assessment VideoLLaMA 2: Advancing Spatial-Temporal Modeling and Audio Understanding in Video-LLMs
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8c990ca2-3e07-4f28-81de-fabe653a10d8 · outbound
AGAV-Rater: Adapting Large Multimodal Model for AI-Generated Audio-Visual Quality Assessment Simple and controllable music generation
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation b185952f-2392-43f5-a3e2-bff21ac70847 · outbound
AGAV-Rater: Adapting Large Multimodal Model for AI-Generated Audio-Visual Quality Assessment PAM: Prompting Audio-Language Models for Audio Quality Assessment
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5ea53f22-e938-4b26-8d79-e8b12116018a · outbound
AGAV-Rater: Adapting Large Multimodal Model for AI-Generated Audio-Visual Quality Assessment Toward universal text-to-music retrieval
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 9aaf1ed9-a3e1-4064-9eb8-702427f8af1f · outbound
AGAV-Rater: Adapting Large Multimodal Model for AI-Generated Audio-Visual Quality Assessment Clap learning audio concepts from natural language supervision
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 7b3f034a-05a7-460c-97fd-1be3d9d0069f · outbound
AGAV-Rater: Adapting Large Multimodal Model for AI-Generated Audio-Visual Quality Assessment VITA: Towards Open-Source Interactive Omni Multimodal LLM
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f2e25c47-60b7-42a2-a8ee-70597ef18393 · outbound
AGAV-Rater: Adapting Large Multimodal Model for AI-Generated Audio-Visual Quality Assessment Gotta Hear Them All: Towards Sound Source Aware Audio Generation
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b1a8fd6b-f6cf-41f4-a5d7-5944ee8d1e82 · outbound
AGAV-Rater: Adapting Large Multimodal Model for AI-Generated Audio-Visual Quality Assessment AnimateDiff : Animate your personalized text-to-image diffusion models without specific tuning
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 95381a4c-0dce-46b3-b59e-285693eb614e · outbound
AGAV-Rater: Adapting Large Multimodal Model for AI-Generated Audio-Visual Quality Assessment Onellm: One framework to align all modalities with language
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 9cc67da7-92c8-46f2-a49a-51bac96a791f · outbound
AGAV-Rater: Adapting Large Multimodal Model for AI-Generated Audio-Visual Quality Assessment Unresolved cited work
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation cefd2335-c222-4665-a4ee-3d47f23f13af · outbound
AGAV-Rater: Adapting Large Multimodal Model for AI-Generated Audio-Visual Quality Assessment CogVideo: Large-scale Pretraining for Text-to-Video Generation via Transformers
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 53accb63-7a54-44d1-bfd5-4d6a6c717ec7 · outbound
AGAV-Rater: Adapting Large Multimodal Model for AI-Generated Audio-Visual Quality Assessment Video-to-Audio Generation with Fine-grained Temporal Semantics
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5d45ad44-df55-4936-996f-d2a5faf2cf52 · outbound
AGAV-Rater: Adapting Large Multimodal Model for AI-Generated Audio-Visual Quality Assessment Vbench: Comprehensive benchmark suite for video generative models
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 073488ba-5728-47ff-9e5d-05e5cf7d8ba3 · outbound
AGAV-Rater: Adapting Large Multimodal Model for AI-Generated Audio-Visual Quality Assessment GPT-4o System Card
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e6f7b572-5f85-4443-939e-3414a4ad8fa8 · outbound
AGAV-Rater: Adapting Large Multimodal Model for AI-Generated Audio-Visual Quality Assessment Taming Visually Guided Sound Generation
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ad621faf-ae9f-4405-9a81-1df5f46d6ee6 · outbound
AGAV-Rater: Adapting Large Multimodal Model for AI-Generated Audio-Visual Quality Assessment Read, watch and scream! sound generation from text and video
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 76f07674-9edd-45ae-86cc-f7209d1b0187 · outbound
AGAV-Rater: Adapting Large Multimodal Model for AI-Generated Audio-Visual Quality Assessment T2VBench : Benchmarking temporal dynamics for text-to-video generation
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 405c95da-158c-48a6-90e4-7c1e87db9807 · outbound
AGAV-Rater: Adapting Large Multimodal Model for AI-Generated Audio-Visual Quality Assessment D., Kim, B., Lee, H., and Kim, G
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation bc894573-ea6a-4ad9-8454-bd76e50531c5 · outbound
AGAV-Rater: Adapting Large Multimodal Model for AI-Generated Audio-Visual Quality Assessment Subjective-aligned dataset and metric for text-to-video quality assessment
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 0423f0dc-9b3c-48f9-9046-375a6fed13ca · outbound
AGAV-Rater: Adapting Large Multimodal Model for AI-Generated Audio-Visual Quality Assessment AudioGen: Textually Guided Audio Generation
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 71c8e2b8-dd21-450b-910e-5d819958bc9a · outbound
AGAV-Rater: Adapting Large Multimodal Model for AI-Generated Audio-Visual Quality Assessment Groundinggpt: Language enhanced multi-modal grounding model
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 25eba396-9234-4e1d-b685-16c12c5f394d · outbound
AGAV-Rater: Adapting Large Multimodal Model for AI-Generated Audio-Visual Quality Assessment VALOR : Vision-audio-language omni-perception pretraining model and dataset
Reference 34
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation d14bee3c-402e-427f-820d-a7e57fb52728 · outbound
AGAV-Rater: Adapting Large Multimodal Model for AI-Generated Audio-Visual Quality Assessment Fetv: A benchmark for fine-grained evaluation of open-domain text-to-video generation
Reference 35
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 4f97cb71-53a0-442b-ae76-0286d41da35d · outbound
AGAV-Rater: Adapting Large Multimodal Model for AI-Generated Audio-Visual Quality Assessment Mosnet: Deep learning based objective assessment for voice conversion
Reference 36
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation e32b041a-0311-420d-8a11-f4ef76a4ad8f · outbound
AGAV-Rater: Adapting Large Multimodal Model for AI-Generated Audio-Visual Quality Assessment Unified-IO 2: Scaling autoregressive multimodal models with vision language audio and action
Reference 37
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 9de5644b-981c-4461-9eb9-988615a2d865 · outbound
AGAV-Rater: Adapting Large Multimodal Model for AI-Generated Audio-Visual Quality Assessment Diff-foley: Synchronized video-to-audio synthesis with latent diffusion models
Reference 38
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 781d5118-91c8-4bd0-859c-7de8313a11cf · outbound
AGAV-Rater: Adapting Large Multimodal Model for AI-Generated Audio-Visual Quality Assessment Speech Quality Assessment through MOS using Non-Matching References
Reference 39
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 87d68138-c513-4a26-8ed8-011dae96b6dc · outbound
AGAV-Rater: Adapting Large Multimodal Model for AI-Generated Audio-Visual Quality Assessment C., and Bovik, A
Reference 40
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation f69d7519-19da-4049-a174-c2f0afd16265 · outbound
AGAV-Rater: Adapting Large Multimodal Model for AI-Generated Audio-Visual Quality Assessment Nisqa: A deep cnn-self-attention model for multidimensional speech quality prediction with crowdsourced datasets
Reference 41
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation a318e777-902e-41cd-a38b-63a683fa931b · outbound
AGAV-Rater: Adapting Large Multimodal Model for AI-Generated Audio-Visual Quality Assessment Audio-visual instance discrimination with cross-modal agreement
Reference 42
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 0173fc3d-43ff-461d-8c7a-f56bc89cdaf7 · outbound
AGAV-Rater: Adapting Large Multimodal Model for AI-Generated Audio-Visual Quality Assessment STA-V2A: Video-to-Audio Generation with Semantic and Temporal Alignment
Reference 43
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 76b4a802-6fbc-4208-b57d-19e5ffa2df38 · outbound
AGAV-Rater: Adapting Large Multimodal Model for AI-Generated Audio-Visual Quality Assessment W., Beerends, J
Reference 44
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation b427e816-91cf-41df-910b-1708a806493d · outbound
AGAV-Rater: Adapting Large Multimodal Model for AI-Generated Audio-Visual Quality Assessment and Adi, Y
Reference 45
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation dcfaa58c-c1fb-4998-ab91-54afefd11800 · outbound
AGAV-Rater: Adapting Large Multimodal Model for AI-Generated Audio-Visual Quality Assessment PandaGPT: One Model To Instruction-Follow Them All
Reference 46
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7d449d2c-7c70-461c-89cd-588cb68bb726 · outbound
AGAV-Rater: Adapting Large Multimodal Model for AI-Generated Audio-Visual Quality Assessment The influence of text-guidance on visual attention
Reference 47
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 120815b1-bc21-40e2-bfd9-36e100e1ba76 · outbound
AGAV-Rater: Adapting Large Multimodal Model for AI-Generated Audio-Visual Quality Assessment How is visual attention influenced by text guidance? database and model
Reference 48
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 8e2c7403-36c5-456d-8351-a3f6fc1535da · outbound
AGAV-Rater: Adapting Large Multimodal Model for AI-Generated Audio-Visual Quality Assessment Explore the Hallucination on Low-level Perception for MLLMs
Reference 49
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation acb85e0e-9cfe-4d1f-8f7e-428eaa8e3022 · outbound
AGAV-Rater: Adapting Large Multimodal Model for AI-Generated Audio-Visual Quality Assessment Gemini 1.5: Unlocking multimodal understanding across millions of tokens of context
Reference 50
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation aa8832c5-e30d-471c-a17d-7504612c1fe3 · outbound
AGAV-Rater: Adapting Large Multimodal Model for AI-Generated Audio-Visual Quality Assessment Reka Core, Flash, and Edge: A Series of Powerful Multimodal Language Models
Reference 51
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation af93185d-f391-436d-bdd6-ecd0f6fea59f · outbound
AGAV-Rater: Adapting Large Multimodal Model for AI-Generated Audio-Visual Quality Assessment V2a-mapper: A lightweight solution for vision-to-audio generation by connecting foundation models
Reference 52
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation a06682c8-17b6-480a-b6c1-0b2897661010 · outbound
AGAV-Rater: Adapting Large Multimodal Model for AI-Generated Audio-Visual Quality Assessment Quality Assessment for AI Generated Images with Instruction Tuning
Reference 53
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cebbe6fc-9da6-49a0-a4a0-2a4688c00092 · outbound
AGAV-Rater: Adapting Large Multimodal Model for AI-Generated Audio-Visual Quality Assessment AIGV-Assessor: Benchmarking and Evaluating the Perceptual Quality of Text-to-Video Generation with LMM
Reference 54
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0a76a232-dce4-41ed-abc9-687eea36cec2 · outbound
AGAV-Rater: Adapting Large Multimodal Model for AI-Generated Audio-Visual Quality Assessment Tiva: Time-aligned video-to-audio generation
Reference 55
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 2fb4dc11-4839-43de-bd19-a2a57c1a7e7a · outbound
AGAV-Rater: Adapting Large Multimodal Model for AI-Generated Audio-Visual Quality Assessment A recipe for scaling up text-to-video generation with text-free videos
Reference 56
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 018796f9-3fd2-43e6-b295-73f652f6c019 · outbound
AGAV-Rater: Adapting Large Multimodal Model for AI-Generated Audio-Visual Quality Assessment Modaverse: Efficiently transforming modalities with llms
Reference 57
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 4e73897b-2810-45a3-94e7-3160415148e0 · outbound
AGAV-Rater: Adapting Large Multimodal Model for AI-Generated Audio-Visual Quality Assessment LAVIE: High-Quality Video Generation with Cascaded Latent Diffusion Models
Reference 58
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3a242a52-48b2-4cec-9a7b-8c4e1b088951 · outbound
AGAV-Rater: Adapting Large Multimodal Model for AI-Generated Audio-Visual Quality Assessment Frieren: Efficient Video-to-Audio Generation Network with Rectified Flow Matching
Reference 59
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 11e0a367-81ea-43cb-9175-79584784ff06 · outbound
AGAV-Rater: Adapting Large Multimodal Model for AI-Generated Audio-Visual Quality Assessment Q-Align : Teaching LMMs for visual scoring via discrete text-defined levels
Reference 60
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation f2864ae5-a3c3-4d5d-806b-0b7be6a7b944 · outbound
AGAV-Rater: Adapting Large Multimodal Model for AI-Generated Audio-Visual Quality Assessment NExT-GPT : Any-to-any multimodal llm
Reference 61
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 50d2b6f4-94a9-42b8-a00c-386de598363e · outbound
AGAV-Rater: Adapting Large Multimodal Model for AI-Generated Audio-Visual Quality Assessment Sonicvisionlm: Playing sound with vision language models
Reference 62
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 910fe5ed-f1b0-48c2-8584-ca20a38da188 · outbound
AGAV-Rater: Adapting Large Multimodal Model for AI-Generated Audio-Visual Quality Assessment Efficient Video to Audio Mapper with Visual Scene Detection
Reference 63
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 795ccb3c-3b8b-435e-840b-12bb25dca64c · outbound
AGAV-Rater: Adapting Large Multimodal Model for AI-Generated Audio-Visual Quality Assessment Telepresence video quality assessment
Reference 64
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 22a6b40c-e1a5-4ea8-8e30-a5d03663809d · outbound
AGAV-Rater: Adapting Large Multimodal Model for AI-Generated Audio-Visual Quality Assessment E., Fu, S.-W., Fuh, C.-S., Tsao, Y., and Wang, H.-M
Reference 65
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 24cb6520-edae-4e6e-ab85-24c89b42df46 · outbound
AGAV-Rater: Adapting Large Multimodal Model for AI-Generated Audio-Visual Quality Assessment FoleyCrafter: Bring Silent Videos to Life with Lifelike and Synchronized Sounds
Reference 66
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0e247d60-34a4-48fc-93ba-a0493d534c6d · outbound
AGAV-Rater: Adapting Large Multimodal Model for AI-Generated Audio-Visual Quality Assessment Lmm-pcqa: Assisting point cloud quality assessment with lmm
Reference 67
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 07557746-ac0d-4d5e-b0b6-243fc6bfc735 · outbound
AGAV-Rater: Adapting Large Multimodal Model for AI-Generated Audio-Visual Quality Assessment write newline
Reference 68
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 13e43b0e-11d7-4144-a940-7bd6e6e57ace · inbound
Efficient Face Image Quality Assessment via Self-training and Knowledge Distillation AGAV-Rater: Adapting Large Multimodal Model for AI-Generated Audio-Visual Quality Assessment
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cfc4fff1-41c7-4d93-badb-65f6addd245b · inbound
Engagement Prediction of Short Videos with Large Multimodal Models AGAV-Rater: Adapting Large Multimodal Model for AI-Generated Audio-Visual Quality Assessment
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.