Pith. sign in

Paper Citation Record · LEDGER

HIPPO-Video: Simulating Watch Histories with Large Language Models for Personalized Video Highlighting

As of 7 August 2026, this Paper Citation Record lists 60 of 60 outbound references and 1 inbound Pith citation observation for arXiv:2507.16873.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2507.16873 v1

Coverage vector

measured 60 of 60 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-06T15:17:58.243690Z

measured 61 of 61 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-05-20T00:55:59.857014Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-20T00:57:54.247973Z

Reference resolution

60 of 60 outbound references displayed

  • verified exact1
  • verified fuzzy40
  • unresolved19
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 111e67ee-b96a-4a46-9964-b9ba7fe484d4 · outbound

This paper cites GPT-4 Technical Report.

HIPPO-Video: Simulating Watch Histories with Large Language Models for Personalized Video Highlighting GPT-4 Technical Report

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-06T15:17:57.694812Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:17:57.694812Z digest=sha256:c811a9edcfa65c7eea335a8ea629a5369cafce35268cf66b5a5385630266945a

Observation 14b28a1e-c8e4-472a-b772-576670cb78ee · outbound

This paper cites Video summarization using deep neural networks: A survey.

HIPPO-Video: Simulating Watch Histories with Large Language Models for Personalized Video Highlighting Video summarization using deep neural networks: A survey

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:18:00.559165Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T15:17:57.706308Z digest=sha256:42976c34731c326d14a6929c49c4cb2d411acfbf62f262f3ddc2442e748d677a

Observation 4f93a273-ade6-4024-b885-94c9a4835eff · outbound

This paper cites Towards automated movie trailer generation.

HIPPO-Video: Simulating Watch Histories with Large Language Models for Personalized Video Highlighting Towards automated movie trailer generation

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:18:00.526219Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T15:17:57.715747Z digest=sha256:b7c806fc7cf4170a9ec3d2eeb57b68cb31baef5935ef092ca5d7fce1f7f1624f

Observation f06a89c7-b8a4-444d-a8e3-b2d4a61b9262 · outbound

This paper cites Scaling up video summarization pretraining with large language models.

HIPPO-Video: Simulating Watch Histories with Large Language Models for Personalized Video Highlighting Scaling up video summarization pretraining with large language models

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:18:00.494069Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T15:17:57.732501Z digest=sha256:1f341c18ff6ebea2dc6de66b36a79f88f8d7ff01a55a53aab050c1cee20e5a21

Observation 9d250bf7-c4b0-4f42-a71a-659e9ba1c1da · outbound

This paper cites Fast-DetectGPT: Efficient Zero-Shot Detection of Machine-Generated Text via Conditional Probability Curvature.

HIPPO-Video: Simulating Watch Histories with Large Language Models for Personalized Video Highlighting Fast-DetectGPT: Efficient Zero-Shot Detection of Machine-Generated Text via Conditional Probability Curvature

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-06T15:17:57.742908Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:17:57.742908Z digest=sha256:e72fa662a3ad4711942ba6a87617ace238c590aa254a94c66ea21d218b090c3f

Observation c5e70f11-e328-4a11-9644-a2a5d94b2764 · outbound

This paper cites End-to-end object detection with transformers.

HIPPO-Video: Simulating Watch Histories with Large Language Models for Personalized Video Highlighting End-to-end object detection with transformers

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-06T15:17:57.754625Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:17:57.754625Z digest=sha256:1d9c925ecac4b31faa736403c131c6f9e29bf95bd5438994c38cddc4edb0fd83

Observation 971a1022-1cae-47db-bb32-8a8d388802f9 · outbound

This paper cites Personalized video summarization by multimodal video understanding.

HIPPO-Video: Simulating Watch Histories with Large Language Models for Personalized Video Highlighting Personalized video summarization by multimodal video understanding

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:18:00.446030Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T15:17:57.772859Z digest=sha256:edf2779a45313048bc12656bbcdf2bf75bd1381692d6407d8655676059e8763a

Observation 2df36854-e189-4c3a-bb60-8d9fccec8895 · outbound

This paper cites Chatbot arena: An open platform for evaluating llms by human preference.

HIPPO-Video: Simulating Watch Histories with Large Language Models for Personalized Video Highlighting Chatbot arena: An open platform for evaluating llms by human preference

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-06T15:17:57.780042Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:17:57.780042Z digest=sha256:048521c381c03e03d127c64003dffc4a4ea69bf3b9a35dd114e10d5f33114c6b

Observation 2f9c3627-2e56-4a5e-84c9-3cce8d9556ee · outbound

This paper cites Tall: Temporal activity localization via language query.

HIPPO-Video: Simulating Watch Histories with Large Language Models for Personalized Video Highlighting Tall: Temporal activity localization via language query

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:18:00.397974Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T15:17:57.794841Z digest=sha256:c267bdacda777a07775accc734a3414ee064ff04c69acb8797b35979600980c1

Observation 6f3e2665-8d99-4161-b88d-a1bbf8e26051 · outbound

This paper cites Creating summaries from user videos.

HIPPO-Video: Simulating Watch Histories with Large Language Models for Personalized Video Highlighting Creating summaries from user videos

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:18:00.361373Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T15:17:57.805901Z digest=sha256:dc689afc294acaeff86e0042a896d876a238b9124d31df42af00e79595ed2edd

Observation 2e82b1b4-5224-40cd-9210-3ccff6cd019a · outbound

This paper cites Video2gif: Automatic generation of animated gifs from video.

HIPPO-Video: Simulating Watch Histories with Large Language Models for Personalized Video Highlighting Video2gif: Automatic generation of animated gifs from video

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:18:00.314682Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T15:17:57.813044Z digest=sha256:7521a125a558353d3ace1da74175ba07928cc1db32d9783b91cdd7eff996fbd8

Observation 6b864fc9-9403-4a05-a35b-e0262781c528 · outbound

This paper cites Shot2Story: A New Benchmark for Comprehensive Understanding of Multi-shot Videos.

HIPPO-Video: Simulating Watch Histories with Large Language Models for Personalized Video Highlighting Shot2Story: A New Benchmark for Comprehensive Understanding of Multi-shot Videos

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-06T15:17:57.824025Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:17:57.824025Z digest=sha256:5e313eddede195c0a4c997a62499c65e6c4de6e98eb9b982ae0c651f37bd18c4

Observation 5b936faf-0f4c-404c-995e-7bdbeb775617 · outbound

This paper cites V2xum-llm: Cross-modal video summarization with temporal prompt instruction tuning.

HIPPO-Video: Simulating Watch Histories with Large Language Models for Personalized Video Highlighting V2xum-llm: Cross-modal video summarization with temporal prompt instruction tuning

Reference 13

Resolution
verified exact
raw_fallback, observed 2026-08-06T15:17:58.535719Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T15:17:57.833243Z digest=sha256:8069e0b760a6a71da5ecd16256c8657a60619fec6c04268fafa37d23aca40a4b

Observation 178f61e5-5ca4-4ccb-bd6a-6b78c0c7cf7e · outbound

This paper cites Movienet: A holistic dataset for movie understanding.

HIPPO-Video: Simulating Watch Histories with Large Language Models for Personalized Video Highlighting Movienet: A holistic dataset for movie understanding

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:18:00.274641Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T15:17:57.850713Z digest=sha256:82f55c9e959c0f235d130d59701e3d72ebfd65a68a0fd3b33cadd706393cefe4

Observation 100bc01e-0ebb-4d22-a71c-1079ae8835d8 · outbound

This paper cites Video summarization with attention-based encoder--decoder networks.

HIPPO-Video: Simulating Watch Histories with Large Language Models for Personalized Video Highlighting Video summarization with attention-based encoder--decoder networks

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:18:00.233683Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T15:17:57.866956Z digest=sha256:774f3d447da7a040f847955086d609dc77b9e50027fbd3fd5f120df2c13bd5ce

Observation b5e92f8f-3ef2-4c14-a09d-c8b297af2f97 · outbound

This paper cites Mdetr-modulated detection for end-to-end multi-modal understanding.

HIPPO-Video: Simulating Watch Histories with Large Language Models for Personalized Video Highlighting Mdetr-modulated detection for end-to-end multi-modal understanding

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-06T15:17:57.878808Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:17:57.878808Z digest=sha256:55a374da2ffe0ebc6ae4e49883d68f4ef60d92ef7eebe05827aaeaa2a9e34014

Observation 78658a8a-e7bf-41e2-9f1d-5f2293280a20 · outbound

This paper cites Self-attentive sequential recommendation.

HIPPO-Video: Simulating Watch Histories with Large Language Models for Personalized Video Highlighting Self-attentive sequential recommendation

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-06T15:17:57.889332Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:17:57.889332Z digest=sha256:cf3061480bcbe24f650f46817061db1fe856473f3656237d4161a6d198cadb4f

Observation 34c922f3-384c-43b9-822b-ac4b3f72e2e4 · outbound

This paper cites Tvr: A large-scale dataset for video-subtitle moment retrieval.

HIPPO-Video: Simulating Watch Histories with Large Language Models for Personalized Video Highlighting Tvr: A large-scale dataset for video-subtitle moment retrieval

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:18:00.099995Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T15:17:57.894979Z digest=sha256:9dcf657407a3651210f2125c85666bf5bab205c715f3e31970351aa886b836dc

Observation 92b3b6c7-adb3-4da3-8832-cbbe029447bf · outbound

This paper cites Detecting moments and highlights in videos via natural language queries.

HIPPO-Video: Simulating Watch Histories with Large Language Models for Personalized Video Highlighting Detecting moments and highlights in videos via natural language queries

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-06T15:17:57.900813Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:17:57.900813Z digest=sha256:ca5f900ffaa2a44224bab4c577bb5e0f5bfe33e8900adb185c7c178350f06ad4

Observation 1186ada7-6db8-4df6-9898-174df4cae153 · outbound

This paper cites HERO: Hierarchical Encoder for Video+Language Omni-representation Pre-training.

HIPPO-Video: Simulating Watch Histories with Large Language Models for Personalized Video Highlighting HERO: Hierarchical Encoder for Video+Language Omni-representation Pre-training

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-06T15:17:57.908906Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:17:57.908906Z digest=sha256:22cbaeddc0394f987c9c0b80f3068263b4b286cc292492f5f9fda8e119cd88a7

Observation 2fffdc20-b80c-4b42-ad59-a56ad77212c1 · outbound

This paper cites Univtg: Towards unified video-language temporal grounding.

HIPPO-Video: Simulating Watch Histories with Large Language Models for Personalized Video Highlighting Univtg: Towards unified video-language temporal grounding

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:18:00.005787Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T15:17:57.915019Z digest=sha256:26a6ea5157f288397fc710218b9956218119ffffc641a6ddd6175ed54b29beef

Observation d74afcc3-c1aa-4ae7-93cd-14992b383a2f · outbound

This paper cites Visual instruction tuning.

HIPPO-Video: Simulating Watch Histories with Large Language Models for Personalized Video Highlighting Visual instruction tuning

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-06T15:17:57.923338Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:17:57.923338Z digest=sha256:d2ba7af32044aac11c72c87940059bd0bb63741ce7a3debe68f2bdc029dd9bf2

Observation fbe27845-2ba9-4e8f-afa8-131644f5e416 · outbound

This paper cites Attentive moment retrieval in videos.

HIPPO-Video: Simulating Watch Histories with Large Language Models for Personalized Video Highlighting Attentive moment retrieval in videos

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:17:59.942980Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T15:17:57.930475Z digest=sha256:7ab066cc4dd059698a3d100f841179f8fdf5e0b078f9f7dc01c4e85add4aa8f4

Observation cb57efb1-20af-4de0-9b32-4c036c771603 · outbound

This paper cites Multi-task deep visual-semantic embedding for video thumbnail selection.

HIPPO-Video: Simulating Watch Histories with Large Language Models for Personalized Video Highlighting Multi-task deep visual-semantic embedding for video thumbnail selection

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:17:59.916236Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T15:17:57.938843Z digest=sha256:a2e4cb00cec839c677910a94116cf5fe96ea6925f469a869904d349e4cbacbdf

Observation 200f224b-e477-4969-b44a-eda34a757df2 · outbound

This paper cites Umt: Unified multi-modal transformers for joint video moment retrieval and highlight detection.

HIPPO-Video: Simulating Watch Histories with Large Language Models for Personalized Video Highlighting Umt: Unified multi-modal transformers for joint video moment retrieval and highlight detection

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:17:59.894062Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T15:17:57.950702Z digest=sha256:10ef1436de883c9be1c9c938004f711e29d0d74285392e71ae0c407674981dd5

Observation 67d4abdc-b013-481e-acd5-af011be4bf6e · outbound

This paper cites Debug: A dense bottom-up grounding approach for natural language video localization.

HIPPO-Video: Simulating Watch Histories with Large Language Models for Personalized Video Highlighting Debug: A dense bottom-up grounding approach for natural language video localization

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:17:59.858277Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T15:17:57.959213Z digest=sha256:b91b3d1da9e87e8fc81e5afb21af10c4a19a73c0f9ceb5670880dfdac78035e0

Observation 14265216-7375-4778-b9c6-da0ad1b99330 · outbound

This paper cites Videoautoarena: An automated arena for evaluating large multimodal models in video analysis through user simulation.

HIPPO-Video: Simulating Watch Histories with Large Language Models for Personalized Video Highlighting Videoautoarena: An automated arena for evaluating large multimodal models in video analysis through user simulation

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:17:59.813663Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T15:17:57.968763Z digest=sha256:03714d8005bfbefed3d8788255bbb560c74213b30b7bd435bdaaea6ead4df279

Observation 6370adc4-c383-4e9f-aa75-86741df5413c · outbound

This paper cites Howto100m: Learning a text-video embedding by watching hundred million narrated video clips.

HIPPO-Video: Simulating Watch Histories with Large Language Models for Personalized Video Highlighting Howto100m: Learning a text-video embedding by watching hundred million narrated video clips

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-06T15:17:57.978108Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:17:57.978108Z digest=sha256:e34ab5e514af35e498208e57905a2d12d35b79856ec9973b0dd719efedc8610c

Observation 13ca2bd9-7318-45af-a15f-56f231f4b646 · outbound

This paper cites Detectgpt: Zero-shot machine-generated text detection using probability curvature.

HIPPO-Video: Simulating Watch Histories with Large Language Models for Personalized Video Highlighting Detectgpt: Zero-shot machine-generated text detection using probability curvature

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:17:59.747306Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T15:17:57.986161Z digest=sha256:1e32ca620ce797e651ebbf091095f78628eda926eacc0022d0d3d9fc0be7b9c8

Observation 12fcf8e1-6451-4820-ac18-dff6ed7fb08f · outbound

This paper cites Query-dependent video representation for moment retrieval and highlight detection.

HIPPO-Video: Simulating Watch Histories with Large Language Models for Personalized Video Highlighting Query-dependent video representation for moment retrieval and highlight detection

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:17:59.706354Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T15:17:57.997352Z digest=sha256:fb3e3ec5d62f935761ee4148f37b846e8f272462d0eb2829682bf2e14f5ca642

Observation cf66e6ea-cd8a-4ed0-b290-2fcea3507307 · outbound

This paper cites Clip-it! language-guided video summarization.

HIPPO-Video: Simulating Watch Histories with Large Language Models for Personalized Video Highlighting Clip-it! language-guided video summarization

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:17:59.673686Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T15:17:58.005873Z digest=sha256:2b37a34257b307acb363a4bcee721378430fe40b1fc3d98b2a0d14ea049fe54d

Observation 69e54f2b-58c6-4089-b36f-483ebc85c181 · outbound

This paper cites Sumgraph: Video summarization via recursive graph modeling.

HIPPO-Video: Simulating Watch Histories with Large Language Models for Personalized Video Highlighting Sumgraph: Video summarization via recursive graph modeling

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:17:59.632181Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T15:17:58.011349Z digest=sha256:8e4cbf806eb725c63c4457d5f3ca2a151158bd59a7d8215f9be996ec92566e86

Observation 0fe70fca-d3e8-468f-bda8-39f3cd0bbfde · outbound

This paper cites Mmsum: A dataset for multimodal summarization and thumbnail generation of videos.

HIPPO-Video: Simulating Watch Histories with Large Language Models for Personalized Video Highlighting Mmsum: A dataset for multimodal summarization and thumbnail generation of videos

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:17:59.587725Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T15:17:58.017617Z digest=sha256:fa66f60789c0b58d4138ced8db50b91c1ece03843571697fe78137a7c7a4d30c

Observation 612b5a05-1947-4417-b6d6-ed88046ad67c · outbound

This paper cites Learning transferable visual models from natural language supervision.

HIPPO-Video: Simulating Watch Histories with Large Language Models for Personalized Video Highlighting Learning transferable visual models from natural language supervision

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-06T15:17:58.026255Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:17:58.026255Z digest=sha256:e9db457834a5fe1a2127ffb05241562d61cdf3f1f539a5667fde9a11b61602d9

Observation 12a5c193-1394-4a55-9fd9-3744ba6d020d · outbound

This paper cites Bpr: Bayesian personalized ranking from implicit feedback.

HIPPO-Video: Simulating Watch Histories with Large Language Models for Personalized Video Highlighting Bpr: Bayesian personalized ranking from implicit feedback

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:17:59.508037Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T15:17:58.034667Z digest=sha256:643f35ae06ba54993cc4ba728b6a076351d5ea4d17ae7642a3b4fb19482f9383

Observation fbca70f4-3f80-48e4-90c8-484a13ac1b35 · outbound

This paper cites Adaptive video highlight detection by learning from user history.

HIPPO-Video: Simulating Watch Histories with Large Language Models for Personalized Video Highlighting Adaptive video highlight detection by learning from user history

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:17:59.462359Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T15:17:58.042738Z digest=sha256:d75e2af8edd2cc5db0e23a7eef868254be47f0d8244b2562895fc7d932683a66

Observation 778060dd-f944-4ab2-b973-c03a99ae4853 · outbound

This paper cites Query-focused extractive video summarization.

HIPPO-Video: Simulating Watch Histories with Large Language Models for Personalized Video Highlighting Query-focused extractive video summarization

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:17:59.415123Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T15:17:58.048754Z digest=sha256:120df506e0eb577b559e4c4de48f508a885b637db6a669ca58549e547df8f735

Observation 068a8527-2123-4f8e-b6f3-a414943f5124 · outbound

This paper cites Query-focused video summarization: Dataset, evaluation, and a memory network based approach.

HIPPO-Video: Simulating Watch Histories with Large Language Models for Personalized Video Highlighting Query-focused video summarization: Dataset, evaluation, and a memory network based approach

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:17:59.379996Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T15:17:58.054036Z digest=sha256:c8f4779c88335ba2a0503c5e75cbb7449042d1883752b87af2feb242b0f7de32

Observation 15bbd598-720d-4cd9-87df-175cf218596e · outbound

This paper cites Tvsum: Summarizing web videos using titles.

HIPPO-Video: Simulating Watch Histories with Large Language Models for Personalized Video Highlighting Tvsum: Summarizing web videos using titles

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:17:59.342002Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T15:17:58.061056Z digest=sha256:41fbd0a4b89371438dfe109dca67ac6a07a971202eeed4a73172465c51096e83

Observation fcf77a06-11a7-4cb3-8e6f-e4e37d409487 · outbound

This paper cites To click or not to click: Automatic selection of beautiful thumbnails from videos.

HIPPO-Video: Simulating Watch Histories with Large Language Models for Personalized Video Highlighting To click or not to click: Automatic selection of beautiful thumbnails from videos

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:17:59.312544Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T15:17:58.067085Z digest=sha256:a80080f80195fe187d736a3d2f03b48ac4e294af37276693dfa4de9a2da584ba

Observation a6087fab-8d80-46c5-90a8-546fbb73ef42 · outbound

This paper cites an unresolved cited work.

HIPPO-Video: Simulating Watch Histories with Large Language Models for Personalized Video Highlighting Unresolved cited work

Reference 41

Resolution
unresolved
raw_fallback, observed 2026-08-06T15:17:59.284673Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T15:17:58.076964Z digest=sha256:f7a0ba1f1a4ea11b9b6cf8fd5b6a0bfc02287e0d03e78b85277e377b1818a239

Observation 1600e304-033b-4d22-a5a0-ccfea5bc1e24 · outbound

This paper cites Tr-detr: Task-reciprocal transformer for joint moment retrieval and highlight detection.

HIPPO-Video: Simulating Watch Histories with Large Language Models for Personalized Video Highlighting Tr-detr: Task-reciprocal transformer for joint moment retrieval and highlight detection

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:17:59.250454Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T15:17:58.087219Z digest=sha256:d19cbfdd41cb7975786cf8d39b0597079dd118018966b243c0e1204b9726366b

Observation 4fedb22b-42da-4fab-a4e1-2b9bc923309d · outbound

This paper cites Ranking domain-specific highlights by analyzing edited videos.

HIPPO-Video: Simulating Watch Histories with Large Language Models for Personalized Video Highlighting Ranking domain-specific highlights by analyzing edited videos

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:17:59.215237Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T15:17:58.099601Z digest=sha256:0b5beb2175a8cfc2289947295965a003887c9173b5aee1f7149fbc053d42a4c1

Observation 87556897-de50-49d6-ba8a-993fb85164d3 · outbound

This paper cites Query-adaptive video summarization via quality-aware relevance estimation.

HIPPO-Video: Simulating Watch Histories with Large Language Models for Personalized Video Highlighting Query-adaptive video summarization via quality-aware relevance estimation

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:17:59.175888Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T15:17:58.108373Z digest=sha256:b445edeb224dbedb7a7466cae4eec209d0b23285a9e6f8a351d09a0b14bfe20e

Observation 11843b59-aebc-42d7-a3f9-d711aec94c8c · outbound

This paper cites Videoagent: Long-form video understanding with large language model as agent.

HIPPO-Video: Simulating Watch Histories with Large Language Models for Personalized Video Highlighting Videoagent: Long-form video understanding with large language model as agent

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-06T15:17:58.113681Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:17:58.113681Z digest=sha256:4d4d779d31d7fe16f3b84a40b9f6ad55ae7526ebeca41fa0b09ffa68523d0753

Observation 5b6a1701-91b5-4a9e-a64b-4b4a2456f62c · outbound

This paper cites Query-biased self-attentive network for query-focused video summarization.

HIPPO-Video: Simulating Watch Histories with Large Language Models for Personalized Video Highlighting Query-biased self-attentive network for query-focused video summarization

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:17:59.096946Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T15:17:58.119440Z digest=sha256:14aec66eff63168573772548608f384d57b25bdf51dcf6eeceef21fa1f054106

Observation 4b5ea4cb-bc56-48d0-b451-09a8d639da2e · outbound

This paper cites Convolutional hierarchical attention network for query-focused video summarization.

HIPPO-Video: Simulating Watch Histories with Large Language Models for Personalized Video Highlighting Convolutional hierarchical attention network for query-focused video summarization

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:17:59.053131Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T15:17:58.126430Z digest=sha256:4182f425b68eae9e63bbcd790d61d3c6427b1d758abc6a4e3554cea8c08dae67

Observation fdfec61e-997d-4624-8fe8-0aaf263c8ae3 · outbound

This paper cites Bridging the gap: A unified video comprehension framework for moment retrieval and highlight detection.

HIPPO-Video: Simulating Watch Histories with Large Language Models for Personalized Video Highlighting Bridging the gap: A unified video comprehension framework for moment retrieval and highlight detection

Reference 48

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:17:59.017479Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T15:17:58.135930Z digest=sha256:e35ba5ab137bfe78ef51a3eb7b30dc4110a3c292e687e3873536ccef2ae2e9fa

Observation a63772b5-cfef-4cb1-8124-3312178cc852 · outbound

This paper cites Cross-category video highlight detection via set-based learning.

HIPPO-Video: Simulating Watch Histories with Large Language Models for Personalized Video Highlighting Cross-category video highlight detection via set-based learning

Reference 49

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:17:58.977655Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T15:17:58.142535Z digest=sha256:eb4c2ce9b2a1af86a2cdcb9dd0809c024bb308a717a2c6982d7361ec75c57c6a

Observation 150ced57-45b7-4bec-a253-c45bab977bf5 · outbound

This paper cites Mh-detr: Video moment and highlight detection with cross-modal transformer.

HIPPO-Video: Simulating Watch Histories with Large Language Models for Personalized Video Highlighting Mh-detr: Video moment and highlight detection with cross-modal transformer

Reference 50

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:17:58.936518Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T15:17:58.147689Z digest=sha256:0c50a8051d4be30e78125a4fe368657b2b22d1a9d0fae1446e46e36245b472ce

Observation ed1d3f71-31c3-44a2-8ad1-67beec4c60e7 · outbound

This paper cites Highlight detection with pairwise deep ranking for first-person video summarization.

HIPPO-Video: Simulating Watch Histories with Large Language Models for Personalized Video Highlighting Highlight detection with pairwise deep ranking for first-person video summarization

Reference 51

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:17:58.909446Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T15:17:58.153170Z digest=sha256:4d1e5919627fbdd43a732cbbc67768e86e94a7d3e8e2bd3fea4d77fc40869fb7

Observation 1326983a-f5e9-476d-af92-e90ef741f5d4 · outbound

This paper cites Semantic conditioned dynamic modulation for temporal sentence grounding in videos.

HIPPO-Video: Simulating Watch Histories with Large Language Models for Personalized Video Highlighting Semantic conditioned dynamic modulation for temporal sentence grounding in videos

Reference 52

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:17:58.872839Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T15:17:58.158832Z digest=sha256:b4ca77dce148c46242efe660992e791aeafd25f1633173ad49d506c5a5a2d6f0

Observation bf2b387c-6436-4f76-a3e4-5da413d9b1f2 · outbound

This paper cites Hierarchical video-moment retrieval and step-captioning.

HIPPO-Video: Simulating Watch Histories with Large Language Models for Personalized Video Highlighting Hierarchical video-moment retrieval and step-captioning

Reference 53

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:17:58.836105Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T15:17:58.165948Z digest=sha256:323837f8430029a1478c810d9a644363ad466e08dc2b6180208a73f2aa0b7468

Observation 04726dc6-2e68-4748-91af-cb19249712e3 · outbound

This paper cites Moment is important: Language-based video moment retrieval via adversarial learning.

HIPPO-Video: Simulating Watch Histories with Large Language Models for Personalized Video Highlighting Moment is important: Language-based video moment retrieval via adversarial learning

Reference 54

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:17:58.809922Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T15:17:58.174479Z digest=sha256:f5da3182bcf49fe5f529c3b2eb76df0b6c332a1ee30557e51e3c2d2e46b5e2d4

Observation 8d2e8ea1-3eb3-4cce-b704-39e7b0c783fe · outbound

This paper cites Span-based Localizing Network for Natural Language Video Localization.

HIPPO-Video: Simulating Watch Histories with Large Language Models for Personalized Video Highlighting Span-based Localizing Network for Natural Language Video Localization

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-06T15:17:58.184421Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:17:58.184421Z digest=sha256:b44ac7293beb5d67bfc566cd6eb4eca223e5d18d8063621cd208b8fb2b764898

Observation 3a048f91-13ce-4e15-b481-71f3410b713f · outbound

This paper cites Towards automatic learning of procedures from web instructional videos.

HIPPO-Video: Simulating Watch Histories with Large Language Models for Personalized Video Highlighting Towards automatic learning of procedures from web instructional videos

Reference 56

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:17:58.779679Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T15:17:58.193137Z digest=sha256:1385a33c1b332bbeadba3cba9f912591413df33df1b33ebd19ceba59795835a1

Observation 8eb3311f-c2fb-43f5-899b-7f6807e76280 · outbound

This paper cites write newline.

HIPPO-Video: Simulating Watch Histories with Large Language Models for Personalized Video Highlighting write newline

Reference 57

Resolution
unresolved
no resolver link, observed 2026-08-06T15:17:58.209470Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:17:58.209470Z digest=sha256:5158d622cd5eb687da7ffb9485f22db2b72812138b32676c4882d5245ed3a474

Observation a8d73be5-de15-45fd-bba3-b628a287f923 · outbound

This paper cites @esa (Ref.

HIPPO-Video: Simulating Watch Histories with Large Language Models for Personalized Video Highlighting @esa (Ref

Reference 58

Resolution
unresolved
no resolver link, observed 2026-08-06T15:17:58.220624Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:17:58.220624Z digest=sha256:90a4e7d153b75096512e9c01e37e236fe86a1858b1856f42580feca5afbbf96b

Observation 221f98c4-a519-4b45-af4a-53619fe51ada · outbound

This paper cites an unresolved cited work.

HIPPO-Video: Simulating Watch Histories with Large Language Models for Personalized Video Highlighting Unresolved cited work

Reference 59

Resolution
unresolved
no resolver link, observed 2026-08-06T15:17:58.235989Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:17:58.235989Z digest=sha256:766a34e94f6f758015c5ddd7c76979cb7b684c8b445a99d21b80c96544df072a

Observation ffbfffb5-013f-445d-90c2-45a901683c77 · outbound

This paper cites an unresolved cited work.

HIPPO-Video: Simulating Watch Histories with Large Language Models for Personalized Video Highlighting Unresolved cited work

Reference 60

Resolution
unresolved
no resolver link, observed 2026-08-06T15:17:58.243690Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:17:58.243690Z digest=sha256:3cc25b302a6765052945b8148b244f18e08a961239837233d834b02ce8ed2757

Pith citing papers

Observation 07909adf-c8b8-4550-8409-6fd6794033a1 · inbound

Will It Go Viral? Grounding Micro-Video Popularity Prediction on the Open Web cites this paper.

Will It Go Viral? Grounding Micro-Video Popularity Prediction on the Open Web HIPPO-Video: Simulating Watch Histories with Large Language Models for Personalized Video Highlighting

Reference 46

Resolution
verified exact
arxiv_id, observed 2026-05-20T00:57:54.249989Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-20T00:55:59.857014Z digest=sha256:a39b76c6fdfb3dbd86c5b84cfc712f0578f64fa3b9ff66de65522201a4e6cb44