Pith. sign in

Paper Citation Record · LEDGER

REGen: Multimodal Retrieval-Embedded Generation for Long-to-Short Video Editing

As of 17 August 2026, this Paper Citation Record lists 41 of 41 outbound references and 0 inbound Pith citation observations for arXiv:2505.18880.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2505.18880 v1

Coverage vector

measured 41 of 41 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T14:29:20.024809Z

measured 41 of 41 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-17T06:30:58.91139+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

41 of 41 outbound references displayed

  • verified exact3
  • verified fuzzy19
  • unresolved17
  • parse uncertain0
  • malformed identifier1
  • metadata mismatch1

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation e283872b-5db8-4bad-941b-ec7e50502c26 · outbound

This paper cites Unveiling the Impact of Multi-Modal Interactions on User Engagement: A Comprehensive Evaluation in AI-driven Conversations.

REGen: Multimodal Retrieval-Embedded Generation for Long-to-Short Video Editing Unveiling the Impact of Multi-Modal Interactions on User Engagement: A Comprehensive Evaluation in AI-driven Conversations

Reference 1

Resolution
verified exact
local_arxiv, observed 2026-08-07T14:29:21.503164Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T14:29:13.821285Z digest=sha256:fc77909e03115f549b9098772be3f886b5f7aa21d9aa1d6ef9c4f26d844b0645

Observation 0ad4a6d7-0ff3-401e-96a4-6481f86c5cd2 · outbound

This paper cites Eye tracking research on readers’ interactions with multimodal texts: a mini-review,.

REGen: Multimodal Retrieval-Embedded Generation for Long-to-Short Video Editing Eye tracking research on readers’ interactions with multimodal texts: a mini-review,

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:29:26.679052Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T14:29:13.916042Z digest=sha256:4efe946dc022d8beda70477a74e1b990fa60fa2f96720b5e186a4d978c6176a2

Observation 0433936a-d9d1-4d8a-a280-72766676c584 · outbound

This paper cites Rec-GPT4V: Multimodal Recommendation with Large Vision-Language Models.

REGen: Multimodal Retrieval-Embedded Generation for Long-to-Short Video Editing Rec-GPT4V: Multimodal Recommendation with Large Vision-Language Models

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-07T14:29:14.045791Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:29:14.045791Z digest=sha256:69373b1077c0d1a863dedcd03cb2b9bd174b5d2b28b0a91a11d45dbe20234239

Observation 1db1b61d-28ea-4596-85e3-07d8a510e747 · outbound

This paper cites Align and Attend: Multimodal Summarization with Dual Contrastive Losses.

REGen: Multimodal Retrieval-Embedded Generation for Long-to-Short Video Editing Align and Attend: Multimodal Summarization with Dual Contrastive Losses

Reference 4

Resolution
verified exact
local_arxiv, observed 2026-08-07T14:29:21.304387Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T14:29:14.172804Z digest=sha256:a00586e7b2a95ae93d6b5785a92069f06a8e732a9e4984198d6cdb8b2cafef59

Observation 00ce0d9d-983b-4e37-a3ca-52c7d07980f0 · outbound

This paper cites Plots to previews: Towards automatic movie preview retrieval using publicly available meta-data,.

REGen: Multimodal Retrieval-Embedded Generation for Long-to-Short Video Editing Plots to previews: Towards automatic movie preview retrieval using publicly available meta-data,

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:29:26.438855Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T14:29:14.351016Z digest=sha256:790552a1b9933f13c9ff00e684c444d42e40abce7c382377d4d5f467c31a13fe

Observation 13544045-6944-48bc-9405-b1a1e51e3b48 · outbound

This paper cites Smart trailer: Automatic generation of movie trailer using only subtitles,.

REGen: Multimodal Retrieval-Embedded Generation for Long-to-Short Video Editing Smart trailer: Automatic generation of movie trailer using only subtitles,

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:29:26.222279Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T14:29:14.531827Z digest=sha256:e2724c09c08c550fa8a2b3f3c1d1a4f9ab1283406a0fdea16fce8e9ab5b19b47

Observation eed9ce76-bd70-4e33-9c35-c2fa8e4ffea9 · outbound

This paper cites Automated production of tv program trailer using electronic program guide,.

REGen: Multimodal Retrieval-Embedded Generation for Long-to-Short Video Editing Automated production of tv program trailer using electronic program guide,

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:29:25.949360Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T14:29:14.693565Z digest=sha256:72cc8bb7fc4ece4f8aa25f9cba982d5af0cc8c31d2d0351e56bc3c8f2c85b521

Observation 197d26a8-09e8-4f18-b029-2e3352eb8a73 · outbound

This paper cites User preferences for automated curation of snackable content,.

REGen: Multimodal Retrieval-Embedded Generation for Long-to-Short Video Editing User preferences for automated curation of snackable content,

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:29:25.704482Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T14:29:14.865046Z digest=sha256:181aec273cdd59a70a2d7623e9cfabc40c953a4f77d33ab23239475ccf2d4929

Observation b9a03670-b177-4baf-b368-8fedc86c964c · outbound

This paper cites Semi-supervised learning towards computerized generation of movie trailers,.

REGen: Multimodal Retrieval-Embedded Generation for Long-to-Short Video Editing Semi-supervised learning towards computerized generation of movie trailers,

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:29:25.364440Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T14:29:14.987099Z digest=sha256:69e69a72d40d4148689a7e747c9398cb28a82f1526272c81235875ccd2755eef

Observation 4b8c28ea-9b35-487b-85a2-d777b4236464 · outbound

This paper cites Harnessing ai for augmenting creativity: Application to movie trailer creation,.

REGen: Multimodal Retrieval-Embedded Generation for Long-to-Short Video Editing Harnessing ai for augmenting creativity: Application to movie trailer creation,

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:29:25.029017Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T14:29:15.178708Z digest=sha256:00de9e62c4d5acf3c652fcff1f3d2ae829a8e3526faac83a2e1ec3dc80d87f1d

Observation aca99c3f-e2a1-49bc-af14-dc0ae6b84918 · outbound

This paper cites TeaserGen: Generating Teasers for Long Documentaries.

REGen: Multimodal Retrieval-Embedded Generation for Long-to-Short Video Editing TeaserGen: Generating Teasers for Long Documentaries

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-07T14:29:15.410134Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:29:15.410134Z digest=sha256:c8ff5e7985bb44db36220c941a3c05accf85357e40c78bfe4a13b8fb28d2eead

Observation fe89d4ab-7b8b-4ef2-85f4-91b98e647d97 · outbound

This paper cites One-Minute Video Generation with Test-Time Training.

REGen: Multimodal Retrieval-Embedded Generation for Long-to-Short Video Editing One-Minute Video Generation with Test-Time Training

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-07T14:29:15.602226Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:29:15.602226Z digest=sha256:f9693c2b7111e0a4b969c55ca8afc9dadfd9c05d0bab636b4d62cff4b6b5bd79

Observation e553459f-b397-4015-b369-eef08cd93f86 · outbound

This paper cites Survey of hallucination in natural language generation,.

REGen: Multimodal Retrieval-Embedded Generation for Long-to-Short Video Editing Survey of hallucination in natural language generation,

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-07T14:29:15.802195Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:29:15.802195Z digest=sha256:1c14998abb778284814734b14ab06e82ba9f42275154716df6841add9d9bbcf2

Observation 29897abb-c2d5-43f6-8d74-e31bb57dbd17 · outbound

This paper cites Hallucination of Multimodal Large Language Models: A Survey.

REGen: Multimodal Retrieval-Embedded Generation for Long-to-Short Video Editing Hallucination of Multimodal Large Language Models: A Survey

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-07T14:29:15.936001Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:29:15.936001Z digest=sha256:0e6c44d486947075adb7263f8ee5a5e22fb618afb77e1ba1475c97ecdbf22a34

Observation 5e603a2a-4222-4987-b25f-7b9a4cfa8408 · outbound

This paper cites Hallucination is Inevitable: An Innate Limitation of Large Language Models.

REGen: Multimodal Retrieval-Embedded Generation for Long-to-Short Video Editing Hallucination is Inevitable: An Innate Limitation of Large Language Models

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-07T14:29:16.063811Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:29:16.063811Z digest=sha256:b2b1b1f5c5934a842e324120944135aa585481b5dcd0588d1981daafca0475e2

Observation 1327115d-b2b1-4216-9fde-40745f918d60 · outbound

This paper cites WhisperX: Time-Accurate Speech Transcription of Long-Form Audio.

REGen: Multimodal Retrieval-Embedded Generation for Long-to-Short Video Editing WhisperX: Time-Accurate Speech Transcription of Long-Form Audio

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-07T14:29:16.220352Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:29:16.220352Z digest=sha256:91c204dabbd65733b6afe6aab14f6dcd6611446a3402a3eced04d61e690bb22f

Observation 6e73a2d4-e2f9-4bc2-9450-016286e49751 · outbound

This paper cites Attributed Question Answering: Evaluation and Modeling for Attributed Large Language Models.

REGen: Multimodal Retrieval-Embedded Generation for Long-to-Short Video Editing Attributed Question Answering: Evaluation and Modeling for Attributed Large Language Models

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-07T14:29:16.346157Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:29:16.346157Z digest=sha256:3744d233df643ab657f17483e5ef1f176127b9978654002da7f56d80be9a0149

Observation 5b05f4e7-53e9-411b-b496-e02bacec2437 · outbound

This paper cites Learning fine-grained grounded citations for attributed large language models,.

REGen: Multimodal Retrieval-Embedded Generation for Long-to-Short Video Editing Learning fine-grained grounded citations for attributed large language models,

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:29:24.643144Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T14:29:16.506154Z digest=sha256:7039addd3c763c50edba19d713b24be59dd1c19dc4cf5ba6cd8a508071239048

Observation a993bad8-35ba-41a8-be34-26186ba1df4d · outbound

This paper cites Dense passage retrieval for open-domain question answering,.

REGen: Multimodal Retrieval-Embedded Generation for Long-to-Short Video Editing Dense passage retrieval for open-domain question answering,

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:29:24.306266Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T14:29:16.633785Z digest=sha256:0f0b559fa70e06817a707937b1dae0fce643659b3ca8253d81c31765b65a3cc8

Observation d28dce8e-ab53-42e9-9971-22ecf4a5c47b · outbound

This paper cites Retrieval-augmented generation for knowledge-intensive nlp tasks,.

REGen: Multimodal Retrieval-Embedded Generation for Long-to-Short Video Editing Retrieval-augmented generation for knowledge-intensive nlp tasks,

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:29:23.966223Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T14:29:16.718104Z digest=sha256:102e7b1be24959c6afc49ec879760ab5cf63017844d5c1f5b87103ccc0581aee

Observation 0b85dbfe-06a0-4221-a98a-74327f2cff90 · outbound

This paper cites Trends in Integration of Knowledge and Large Language Models: A Survey and Taxonomy of Methods, Benchmarks, and Applications.

REGen: Multimodal Retrieval-Embedded Generation for Long-to-Short Video Editing Trends in Integration of Knowledge and Large Language Models: A Survey and Taxonomy of Methods, Benchmarks, and Applications

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-07T14:29:16.912491Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:29:16.912491Z digest=sha256:dec87d0e73651012985937e1096fb1d0dc0624d42025cd39934f1e7e49651ac7

Observation 8261d55b-cc8e-4d76-b4cf-386f27004647 · outbound

This paper cites LaMDA: Language Models for Dialog Applications.

REGen: Multimodal Retrieval-Embedded Generation for Long-to-Short Video Editing LaMDA: Language Models for Dialog Applications

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-07T14:29:17.043717Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:29:17.043717Z digest=sha256:d291068a37e27172a063c6d763ca2b9392dbeca300608239f717bc8cd9e1b588

Observation d94cd4b9-6c70-47ac-8b0d-943d15a76c7a · outbound

This paper cites Evaluating verifiability in generative search engines,.

REGen: Multimodal Retrieval-Embedded Generation for Long-to-Short Video Editing Evaluating verifiability in generative search engines,

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:29:23.618273Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T14:29:17.243963Z digest=sha256:c3f1fa29a62cc6e34c6076aa84c9a684fbd282ed6d8a95e5905f0c770a63360c

Observation 6cbb0576-5b99-4f9c-bec4-1132ffa6d156 · outbound

This paper cites Clip-it! language-guided video summarization,.

REGen: Multimodal Retrieval-Embedded Generation for Long-to-Short Video Editing Clip-it! language-guided video summarization,

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:29:23.296737Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T14:29:17.345158Z digest=sha256:e2013c0bfaa7c103864e177582ea6d5b110f4eeca86d19051a995b7b23312339

Observation 755f5b7f-afd3-40d4-8d6d-1b04e322d7a3 · outbound

This paper cites Scaling Up Video Summarization Pretraining with Large Language Models.

REGen: Multimodal Retrieval-Embedded Generation for Long-to-Short Video Editing Scaling Up Video Summarization Pretraining with Large Language Models

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-07T14:29:17.536205Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:29:17.536205Z digest=sha256:f23936899b2c5ec3e4946d535e3efde35470b1323920735cc824bcf36f99ec06

Observation 794680f7-d888-4de9-9749-587bbb685941 · outbound

This paper cites Towards Automated Movie Trailer Generation.

REGen: Multimodal Retrieval-Embedded Generation for Long-to-Short Video Editing Towards Automated Movie Trailer Generation

Reference 26

Resolution
verified exact
local_arxiv, observed 2026-08-07T14:29:20.939408Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T14:29:17.663260Z digest=sha256:a99cfe27f11579cd8d374eb7938c76ac1341e40cfd74d91c0f12649c9cc5debb

Observation 995a39ee-b9a4-46ae-b5a6-42bfa0740cc9 · outbound

This paper cites "previously on.

REGen: Multimodal Retrieval-Embedded Generation for Long-to-Short Video Editing "previously on

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:29:23.046760Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T14:29:17.787983Z digest=sha256:bbd2801b5a61ec13b46fbabf0ce6f52db062b0972439c7ce1cc1f6d9efd66784

Observation b5b07bb0-2b4b-428a-bf4c-084710c66a73 · outbound

This paper cites Videoxum: Cross-modal visual and textural summarization of videos,.

REGen: Multimodal Retrieval-Embedded Generation for Long-to-Short Video Editing Videoxum: Cross-modal visual and textural summarization of videos,

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:29:22.767143Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T14:29:18.143672Z digest=sha256:d861016ffff6f95fb6e2d4d8f6cf87f9d17f6ec8d0dc87792b9f4d120a4a3f49

Observation 8f9a7602-70c7-4fb0-a5ed-075663663f95 · outbound

This paper cites GPT-4 Technical Report.

REGen: Multimodal Retrieval-Embedded Generation for Long-to-Short Video Editing GPT-4 Technical Report

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-07T14:29:18.310362Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:29:18.310362Z digest=sha256:731a3ff73df22ca9afeae80bcd8e7a60022ae8a1d80867bc3150f8d7b3b0875c

Observation 611a5f88-a102-4d55-b1a4-5aa99f99c993 · outbound

This paper cites LLaMA: Open and Efficient Foundation Language Models.

REGen: Multimodal Retrieval-Embedded Generation for Long-to-Short Video Editing LLaMA: Open and Efficient Foundation Language Models

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-07T14:29:18.464070Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:29:18.464070Z digest=sha256:004b3f5d3a9f124cfa03b8bb5239c48f33af072bf9b60a7ecdeff0d51dfd6c7f

Observation 15dfc0bf-8ef2-4f11-b1c5-172cd213cab2 · outbound

This paper cites Univtg: Towards unified video-language temporal grounding,.

REGen: Multimodal Retrieval-Embedded Generation for Long-to-Short Video Editing Univtg: Towards unified video-language temporal grounding,

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:29:22.428707Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T14:29:18.600886Z digest=sha256:082c31210f6e3e0273c24ec4dbf380bc6bf6a10405ae8328f4f667f7fc207aa1

Observation b9fd862c-5ee2-4632-8fa3-04215822810b · outbound

This paper cites BART: Denoising sequence-to-sequence pre-training for natural language generation, translation, and comprehension,.

REGen: Multimodal Retrieval-Embedded Generation for Long-to-Short Video Editing BART: Denoising sequence-to-sequence pre-training for natural language generation, translation, and comprehension,

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:29:22.276219Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T14:29:18.746817Z digest=sha256:59bbc0e30ade8c438dc161dd2230f94827f389ce56b3c2cb840160aee69c606d

Observation a8ac58e6-26a8-49dd-8ca4-04d4c6d8141b · outbound

This paper cites Sentence-Transformers/all-mpnet-base-v2,.

REGen: Multimodal Retrieval-Embedded Generation for Long-to-Short Video Editing Sentence-Transformers/all-mpnet-base-v2,

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:29:22.152397Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T14:29:18.921410Z digest=sha256:e4e3e58f4a3bf5da2d8702075e442e5176449591f256b6cfc3c799c7c815e5e8

Observation 33152d2e-3f97-487e-8a01-c14956feedc9 · outbound

This paper cites ROUGE: A package for automatic evaluation of summaries,.

REGen: Multimodal Retrieval-Embedded Generation for Long-to-Short Video Editing ROUGE: A package for automatic evaluation of summaries,

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-07T14:29:19.091604Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:29:19.091604Z digest=sha256:ba887383958856f44cfb67943810da4d73b30826c050acf06ea2df71fef5fbd4

Observation 5825d89c-5d59-42d6-8a1b-3f7c73a2da59 · outbound

This paper cites G-Eval: NLG Evaluation using GPT-4 with Better Human Alignment.

REGen: Multimodal Retrieval-Embedded Generation for Long-to-Short Video Editing G-Eval: NLG Evaluation using GPT-4 with Better Human Alignment

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-07T14:29:19.209163Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:29:19.209163Z digest=sha256:b7e2303b4f81fdc6ee3cffbbded0022f157636fcd5d61f120bc74da6501ac9fd

Observation 1a23693b-7bbf-4a73-bb6a-a94c48240283 · outbound

This paper cites DeepEval: The open-source llm evaluation framework,.

REGen: Multimodal Retrieval-Embedded Generation for Long-to-Short Video Editing DeepEval: The open-source llm evaluation framework,

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:29:21.909800Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T14:29:19.386401Z digest=sha256:8d7daa8686b5991a0edc7cebcd5af447d6e728ba9f3d91c50eda9ca7305cdd0b

Observation 6c7bcdc1-29a1-4a3e-a923-eb52436eb22f · outbound

This paper cites CLIPScore: A Reference-free Evaluation Metric for Image Captioning.

REGen: Multimodal Retrieval-Embedded Generation for Long-to-Short Video Editing CLIPScore: A Reference-free Evaluation Metric for Image Captioning

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-07T14:29:19.582613Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:29:19.582613Z digest=sha256:7049fa1d95314571bf59e9a7a0d065c3f3ead3486ad6eb0f7c51a8368c725b01

Observation 157f6fda-85e1-463a-a81e-8843022d8fda · outbound

This paper cites Screenwriter: Automatic screenplay generation and movie summarisation,.

REGen: Multimodal Retrieval-Embedded Generation for Long-to-Short Video Editing Screenwriter: Automatic screenplay generation and movie summarisation,

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:29:21.731470Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T14:29:19.738415Z digest=sha256:2a0dc44ba97f1b3104a66745352fb2d827abe2bd9c9004512a05df420afa8f6a

Observation 774a4142-778a-4689-83df-0ffee8a6b303 · outbound

This paper cites Multimodal Lecture Presentations Dataset: Understanding Multimodality in Educational Slides.

REGen: Multimodal Retrieval-Embedded Generation for Long-to-Short Video Editing Multimodal Lecture Presentations Dataset: Understanding Multimodality in Educational Slides

Reference 39

Resolution
malformed identifier
local_arxiv, observed 2026-08-07T14:29:20.283989Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T14:29:20.024809Z digest=sha256:7d0691e5392bec43acfdedfd448f3a8b10e32cab0a80225cc69307addc331be3

Observation 7b99f98a-f4e5-4dfd-9b63-c8b86f0a807a · outbound

This paper cites ScreenWriter: Automatic Screenplay Generation and Movie Summarisation.

REGen: Multimodal Retrieval-Embedded Generation for Long-to-Short Video Editing ScreenWriter: Automatic Screenplay Generation and Movie Summarisation

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-07T14:29:19.875241Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:29:19.875241Z digest=sha256:9c96284413564d40f2adcf4fa905f5153630dd46380a4643bb32b4c52c18a2c4

Observation 2576b8f1-a8db-442e-b5c9-50ee6f3a5698 · outbound

This paper cites "Previously on ..." From Recaps to Story Summarization.

REGen: Multimodal Retrieval-Embedded Generation for Long-to-Short Video Editing "Previously on ..." From Recaps to Story Summarization

Reference 2024

Resolution
metadata mismatch
local_arxiv, observed 2026-08-07T14:29:20.705164Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T14:29:17.974793Z digest=sha256:16a5db19539a6c9413a9f2d9c5fd08f36ce493319f5a79339bc9856857d4f1ee

Pith citing papers

No inbound Pith citation observations are available.