Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-12T21:27:16.884770Z
Paper Citation Record · LEDGER
As of 14 August 2026, this Paper Citation Record lists 100 of 132 outbound references and 2 inbound Pith citation observations for arXiv:2411.08753.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-12T21:27:16.884770Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-14T06:32:32.682623+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-11T04:45:38.441181Z
A source-named dated measurement, never combined with another source.
Source: pith, observed 2026-08-07T06:00:58.843605Z
100 of 132 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation bb025d11-7abc-45fa-941e-be55fa65712e · outbound
Which Viewpoint Shows it Best? Language for Weakly Supervising View Selection in Multi-view Instructional Videos Deep Learning using Rectified Linear Units (ReLU)
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bbbd5838-3012-4a00-bed1-68e5f1a12eb2 · outbound
Which Viewpoint Shows it Best? Language for Weakly Supervising View Selection in Multi-view Instructional Videos McCrae, Kenton Murray, Maria Nadejde, Satoshi Nakamura, Matteo Negri, Ha Nguyen, Jan Niehues, Xing Niu, Atul Kr
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5f7e63c4-d784-42c4-9cd6-80c062bfff39 · outbound
Which Viewpoint Shows it Best? Language for Weakly Supervising View Selection in Multi-view Instructional Videos A dataset for develop- ing and benchmarking active vision
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5f35b769-c0b0-401e-b12b-d158305d7e1d · outbound
Which Viewpoint Shows it Best? Language for Weakly Supervising View Selection in Multi-view Instructional Videos Automatic editing of footage from multi- ple social cameras
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2b809dd6-454e-4425-a264-bae399839d18 · outbound
Which Viewpoint Shows it Best? Language for Weakly Supervising View Selection in Multi-view Instructional Videos Unresolved cited work
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cc8180e4-325e-4e9e-8930-f4a1fede591c · outbound
Which Viewpoint Shows it Best? Language for Weakly Supervising View Selection in Multi-view Instructional Videos WeaQA: Weak Supervision via Captions for Visual Question Answering
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a3137ed3-695d-457c-b997-5a27ba7b661f · outbound
Which Viewpoint Shows it Best? Language for Weakly Supervising View Selection in Multi-view Instructional Videos METEOR: An auto- matic metric for MT evaluation with improved correlation with human judgments
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1fd79043-4e34-4a2c-892b-dd4ab3137967 · outbound
Which Viewpoint Shows it Best? Language for Weakly Supervising View Selection in Multi-view Instructional Videos Is space-time attention all you need for video understanding? In ICML, page 4, 2021
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 71bc575b-f19d-4b46-8a5c-939063e2b0b4 · outbound
Which Viewpoint Shows it Best? Language for Weakly Supervising View Selection in Multi-view Instructional Videos High- lightme: Detecting highlights from human-centric videos
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1970d812-711b-4c45-a19f-73c8b38c4a28 · outbound
Which Viewpoint Shows it Best? Language for Weakly Supervising View Selection in Multi-view Instructional Videos Extreme rotation estimation using dense correlation volumes
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0a0c9b5d-fb6a-4dba-a770-a30c788473ae · outbound
Which Viewpoint Shows it Best? Language for Weakly Supervising View Selection in Multi-view Instructional Videos Davis, and Lei Zhang
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e8215ca4-bb01-45c4-917f-b677ab0198c0 · outbound
Which Viewpoint Shows it Best? Language for Weakly Supervising View Selection in Multi-view Instructional Videos Enhanced interactive 360° viewing via automatic guidance
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 20ccd794-4833-45b6-95bf-6553e0024242 · outbound
Which Viewpoint Shows it Best? Language for Weakly Supervising View Selection in Multi-view Instructional Videos Learn- ing sports camera selection from internet videos
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9f637a90-1a36-4165-9a78-1202d46acbd4 · outbound
Which Viewpoint Shows it Best? Language for Weakly Supervising View Selection in Multi-view Instructional Videos Wide- baseline relative camera pose estimation with directional learning
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3899a749-9f6c-48a5-bf29-582af09d9ade · outbound
Which Viewpoint Shows it Best? Language for Weakly Supervising View Selection in Multi-view Instructional Videos Geometry-aware recurrent neural networks for active visual recognition
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 128e8bf9-6538-4aff-8651-901cd8535971 · outbound
Which Viewpoint Shows it Best? Language for Weakly Supervising View Selection in Multi-view Instructional Videos Towards a richer 2d understanding of hands at scale
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f512be94-a5f9-47ad-9c48-3fe5ebc5064b · outbound
Which Viewpoint Shows it Best? Language for Weakly Supervising View Selection in Multi-view Instructional Videos Gonzalez, Ion Stoica, and Eric P
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 85501816-c5e0-4811-a324-9150f4030214 · outbound
Which Viewpoint Shows it Best? Language for Weakly Supervising View Selection in Multi-view Instructional Videos Self-view Grounding Given a Narrated 360{\deg} Video
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0b7dded2-0479-4e31-860c-2989caae0cf8 · outbound
Which Viewpoint Shows it Best? Language for Weakly Supervising View Selection in Multi-view Instructional Videos Video co-summarization: Video summarization by visual co- occurrence
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7ab7f5a7-5e64-401d-b3e0-1b9d56e9b9a2 · outbound
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 15cdcf51-af90-48d1-9276-bccd1ddac3ad · outbound
Which Viewpoint Shows it Best? Language for Weakly Supervising View Selection in Multi-view Instructional Videos Scaling egocentric vision: The epic- kitchens dataset
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 13f61efd-951b-41cd-a55b-661009310c20 · outbound
Which Viewpoint Shows it Best? Language for Weakly Supervising View Selection in Multi-view Instructional Videos FlashAttention-2: Faster Attention with Better Parallelism and Work Partitioning
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 58020f30-05f1-4628-8f05-3ada0da5fc75 · outbound
Which Viewpoint Shows it Best? Language for Weakly Supervising View Selection in Multi-view Instructional Videos Flashattention: Fast and memory-efficient exact attention with io-awareness
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9cb4724f-b0dd-49f3-825f-fe2f18162725 · outbound
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d0be0b1f-e948-4cb3-b0af-4b86378f439e · outbound
Which Viewpoint Shows it Best? Language for Weakly Supervising View Selection in Multi-view Instructional Videos Virtex: Learning visual representations from textual annotations
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation acde62f6-cf94-455a-bc65-539a3d3f779d · outbound
Which Viewpoint Shows it Best? Language for Weakly Supervising View Selection in Multi-view Instructional Videos An Image is Worth 16x16 Words: Transformers for Image Recognition at Scale
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 786245bb-34cc-4782-b6da-f8e960b5843f · outbound
Which Viewpoint Shows it Best? Language for Weakly Supervising View Selection in Multi-view Instructional Videos Dense and aligned captions (dac) promote compositional reasoning in vl models
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation face0574-4d45-4a9a-b547-52f30f5a2cb4 · outbound
Which Viewpoint Shows it Best? Language for Weakly Supervising View Selection in Multi-view Instructional Videos Multi-view active fine- grained visual recognition
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ae7829db-5a61-4529-b7b2-53d081487132 · outbound
Which Viewpoint Shows it Best? Language for Weakly Supervising View Selection in Multi-view Instructional Videos Multi- stream dynamic video summarization
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6c41b363-bd5c-48fc-8940-827deb8740d4 · outbound
Which Viewpoint Shows it Best? Language for Weakly Supervising View Selection in Multi-view Instructional Videos Elson and Mark O
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 20a10721-ce03-4f24-8f5d-68767cb25750 · outbound
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0904ea70-ee80-4a76-9aac-ec3227472dea · outbound
Which Viewpoint Shows it Best? Language for Weakly Supervising View Selection in Multi-view Instructional Videos Multi-view video summa- rization
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation aea03a6d-8bf6-41c2-823e-ae6aaa5262d3 · outbound
Which Viewpoint Shows it Best? Language for Weakly Supervising View Selection in Multi-view Instructional Videos Gleicher, Rachel M
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 69cd9eef-ad25-4bde-be6c-0f56da3ca3fa · outbound
Which Viewpoint Shows it Best? Language for Weakly Supervising View Selection in Multi-view Instructional Videos PEAVS: Perceptual Evaluation of Audio-Visual Synchrony Grounded in Viewers' Opinion Scores
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b399e38c-c960-45a5-8172-81faef191d2f · outbound
Which Viewpoint Shows it Best? Language for Weakly Supervising View Selection in Multi-view Instructional Videos Diverse sequential subset selection for supervised video summarization
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1fb9a401-425c-4436-beaa-101ae1ca7ace · outbound
Which Viewpoint Shows it Best? Language for Weakly Supervising View Selection in Multi-view Instructional Videos Ego4d: Around the world in 3,000 hours of egocentric video
Reference 36
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c0b2dc9a-8d8c-43ed-9738-c519d6a0c022 · outbound
Which Viewpoint Shows it Best? Language for Weakly Supervising View Selection in Multi-view Instructional Videos Ego-Exo4D: Understanding Skilled Human Activity from First- and Third-Person Perspectives
Reference 37
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 36ada9b3-36b4-4a4b-bdaf-73478b66a597 · outbound
Which Viewpoint Shows it Best? Language for Weakly Supervising View Selection in Multi-view Instructional Videos Temporal Difference Variational Auto-Encoder
Reference 38
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 82cf2f02-60ce-48a7-b2b7-45c7745d63b4 · outbound
Which Viewpoint Shows it Best? Language for Weakly Supervising View Selection in Multi-view Instructional Videos From Images to Textual Prompts: Zero-shot VQA with Frozen Large Language Models
Reference 39
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bb7f5658-868b-4da4-8003-6fea52ca632f · outbound
Which Viewpoint Shows it Best? Language for Weakly Supervising View Selection in Multi-view Instructional Videos Using closed captions as supervision for video activity recognition
Reference 40
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ee3b7321-6456-4c50-a586-d310a09ced60 · outbound
Which Viewpoint Shows it Best? Language for Weakly Supervising View Selection in Multi-view Instructional Videos Creating summaries from user videos
Reference 41
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation adf155bf-20d6-4c5c-b6cc-30895c8ea6ba · outbound
Which Viewpoint Shows it Best? Language for Weakly Supervising View Selection in Multi-view Instructional Videos Video summarization by learning submodular mixtures of objec- tives
Reference 42
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 51c0c39d-eb07-4531-b1e0-32883fb4dd9b · outbound
Which Viewpoint Shows it Best? Language for Weakly Supervising View Selection in Multi-view Instructional Videos Align and attend: Multimodal summarization with dual contrastive losses
Reference 43
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ccfe8115-fc24-4cb6-8957-db3605f10956 · outbound
Which Viewpoint Shows it Best? Language for Weakly Supervising View Selection in Multi-view Instructional Videos Cohen, and David H
Reference 44
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4771a9c2-2ac3-4484-be01-c8bc7857d19c · outbound
Which Viewpoint Shows it Best? Language for Weakly Supervising View Selection in Multi-view Instructional Videos Cohen, and David H
Reference 45
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 772d4bde-00fa-4203-a507-cb95825b6041 · outbound
Which Viewpoint Shows it Best? Language for Weakly Supervising View Selection in Multi-view Instructional Videos Vir- tual videography
Reference 46
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 00f95f17-6312-4f34-91fb-bd507296c3bd · outbound
Which Viewpoint Shows it Best? Language for Weakly Supervising View Selection in Multi-view Instructional Videos LoRA: Low-Rank Adaptation of Large Language Models
Reference 47
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d38383b6-f858-4707-a017-1dede39c461d · outbound
Which Viewpoint Shows it Best? Language for Weakly Supervising View Selection in Multi-view Instructional Videos Deep 360 pilot: Learning a deep agent for piloting through 360deg sports videos
Reference 48
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation eaaccbea-74bd-4aa4-a9a1-0ed54ce64a55 · outbound
Which Viewpoint Shows it Best? Language for Weakly Supervising View Selection in Multi-view Instructional Videos EgoExoLearn: A Dataset for Bridging Asynchronous Ego- and Exo-centric View of Procedural Activities in Real World
Reference 49
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d3df167d-f2ce-44a1-97d9-fcefb64f1699 · outbound
Which Viewpoint Shows it Best? Language for Weakly Supervising View Selection in Multi-view Instructional Videos Batch normalization: accelerating deep network training by reducing internal co- variate shift
Reference 50
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c0397a64-a4a6-40dc-9e1d-f306b0a6255a · outbound
Which Viewpoint Shows it Best? Language for Weakly Supervising View Selection in Multi-view Instructional Videos Look-ahead be- fore you leap: end-to-end active recognition by forecasting the effect of motion
Reference 51
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8ea4e95b-7769-44eb-b8c4-b79cb4233be3 · outbound
Which Viewpoint Shows it Best? Language for Weakly Supervising View Selection in Multi-view Instructional Videos Learning to look around: Intelligently exploring unseen environments for unknown tasks
Reference 52
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 512bb144-e0c5-4bf9-a5f8-977456cfdf53 · outbound
Which Viewpoint Shows it Best? Language for Weakly Supervising View Selection in Multi-view Instructional Videos End-to-end policy learning for active visual categorization
Reference 53
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2835bd81-9f36-41e4-a29d-cbf613bc3efc · outbound
Which Viewpoint Shows it Best? Language for Weakly Supervising View Selection in Multi-view Instructional Videos Time-Agnostic Prediction: Predicting Predictable Video Frames
Reference 54
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5c8021b6-ddd4-4df0-a69b-51c0e2d0c7ce · outbound
Which Viewpoint Shows it Best? Language for Weakly Supervising View Selection in Multi-view Instructional Videos Simglim: Simplifying glimpse based active visual reconstruction
Reference 55
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d3adfba1-3679-4606-90ce-c0a8591e3d1d · outbound
Which Viewpoint Shows it Best? Language for Weakly Supervising View Selection in Multi-view Instructional Videos Lemma: A multi-view dataset for le arning m ulti-agent m ulti-task a ctivities
Reference 56
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 75304319-c3b5-4630-854f-269cd0168e43 · outbound
Which Viewpoint Shows it Best? Language for Weakly Supervising View Selection in Multi-view Instructional Videos RTMPose: Real-Time Multi-Person Pose Estimation based on MMPose
Reference 57
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cece5df7-2619-49f6-be83-94a62c0fa7f0 · outbound
Which Viewpoint Shows it Best? Language for Weakly Supervising View Selection in Multi-view Instructional Videos Large-scale video summarization using web-image priors
Reference 58
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4afe155d-955c-48ed-bbcc-329d56768cd9 · outbound
Which Viewpoint Shows it Best? Language for Weakly Supervising View Selection in Multi-view Instructional Videos Unresolved cited work
Reference 59
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cb661ab1-ee16-4e98-81ab-c87070f9aa7b · outbound
Which Viewpoint Shows it Best? Language for Weakly Supervising View Selection in Multi-view Instructional Videos Segment any- thing
Reference 60
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 883c89c2-d078-4ec9-9dee-f367fe283eba · outbound
Which Viewpoint Shows it Best? Language for Weakly Supervising View Selection in Multi-view Instructional Videos Hyperbolic Learning with Synthetic Captions for Open-World Detection
Reference 61
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 41dfb3fc-a295-4cc4-895d-81fc3c3afa56 · outbound
Which Viewpoint Shows it Best? Language for Weakly Supervising View Selection in Multi-view Instructional Videos A memory network approach for story-based temporal summarization of 360° videos
Reference 62
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f6e45f0f-680a-4afb-9c6e-ea8b2a8a707a · outbound
Which Viewpoint Shows it Best? Language for Weakly Supervising View Selection in Multi-view Instructional Videos Predicting important objects for egocentric video summarization
Reference 63
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ebfefeba-e6cc-4e98-ad04-5e03f150fd7a · outbound
Which Viewpoint Shows it Best? Language for Weakly Supervising View Selection in Multi-view Instructional Videos MVBench: A Comprehensive Multi-modal Video Understanding Benchmark
Reference 64
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e15fd5d7-482f-47e8-9912-453c1c504e1c · outbound
Which Viewpoint Shows it Best? Language for Weakly Supervising View Selection in Multi-view Instructional Videos Grounded language-image pre-training
Reference 65
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cb28cbc6-1078-4f06-883f-5b54aacfb4cb · outbound
Which Viewpoint Shows it Best? Language for Weakly Supervising View Selection in Multi-view Instructional Videos How local is the local diversity? reinforcing sequen- tial determinantal point processes with dynamic ground sets for supervised video summarization
Reference 66
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 756366c4-920a-4fa0-adff-8122bda308df · outbound
Which Viewpoint Shows it Best? Language for Weakly Supervising View Selection in Multi-view Instructional Videos Egocentric video-language pretraining
Reference 67
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation cadd1a59-85f1-4e16-aa5a-d1d422355b18 · outbound
Which Viewpoint Shows it Best? Language for Weakly Supervising View Selection in Multi-view Instructional Videos Grounding DINO: Marrying DINO with Grounded Pre-Training for Open-Set Object Detection
Reference 68
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d116433c-7261-411f-861d-40dd7605c0ca · outbound
Which Viewpoint Shows it Best? Language for Weakly Supervising View Selection in Multi-view Instructional Videos SGDR: Stochastic Gradient Descent with Warm Restarts
Reference 69
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 30c3a2d4-49dd-4644-9a7e-f73abf1a3b68 · outbound
Which Viewpoint Shows it Best? Language for Weakly Supervising View Selection in Multi-view Instructional Videos Decoupled Weight Decay Regularization
Reference 70
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5fefd853-04f0-49ae-bd8c-3265e7192f38 · outbound
Which Viewpoint Shows it Best? Language for Weakly Supervising View Selection in Multi-view Instructional Videos Story-driven summariza- tion for egocentric video
Reference 71
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation a9ae5bc4-44f2-43e1-b13f-e77415b31e02 · outbound
Which Viewpoint Shows it Best? Language for Weakly Supervising View Selection in Multi-view Instructional Videos Switch-a-View: View Selection Learned from Unlabeled In-the-wild Videos
Reference 72
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 04784fba-8090-4375-bf5d-e45497f5a31e · outbound
Which Viewpoint Shows it Best? Language for Weakly Supervising View Selection in Multi-view Instructional Videos Video summarization via multi- view representative selection
Reference 73
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation fe0807fc-36ba-4265-b1fc-4fa9b5e49f6a · outbound
Which Viewpoint Shows it Best? Language for Weakly Supervising View Selection in Multi-view Instructional Videos Howto100m: Learning a text-video embedding by watching hundred million narrated video clips
Reference 74
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 8106cd4f-a3de-45d2-9b5d-29a59d8cef9b · outbound
Which Viewpoint Shows it Best? Language for Weakly Supervising View Selection in Multi-view Instructional Videos Srinivasan, Matthew Tancik, Jonathan T
Reference 75
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 03ae2a1e-6682-4d9f-a1ec-1b14efc1a1e2 · outbound
Which Viewpoint Shows it Best? Language for Weakly Supervising View Selection in Multi-view Instructional Videos Automatized summarization of multi- player games
Reference 76
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 61da99fb-d31c-4453-92b7-c0c58c127885 · outbound
Which Viewpoint Shows it Best? Language for Weakly Supervising View Selection in Multi-view Instructional Videos Egoenv: Human- centric environment representations from egocentric video
Reference 77
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 67c3248e-d510-4503-9440-8e7f5cc3a3bb · outbound
Which Viewpoint Shows it Best? Language for Weakly Supervising View Selection in Multi-view Instructional Videos Tl; dw? summarizing instructional videos with task relevance and cross-modal saliency
Reference 78
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 82e5f804-aefd-4974-b8dc-ef24f38a6c01 · outbound
Which Viewpoint Shows it Best? Language for Weakly Supervising View Selection in Multi-view Instructional Videos Adaptive skip intervals: Temporal abstraction for recurrent dynamical models
Reference 79
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 3e63b266-ec5b-4f6c-abe9-3c919f19deda · outbound
Which Viewpoint Shows it Best? Language for Weakly Supervising View Selection in Multi-view Instructional Videos Au- tomatic video summarization by graph modeling
Reference 80
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation ad9bab32-aae4-4415-8174-2a602a225541 · outbound
Which Viewpoint Shows it Best? Language for Weakly Supervising View Selection in Multi-view Instructional Videos Collabora- tive summarization of topic-related videos
Reference 81
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 1deecac3-7788-4b32-ab77-5d47a72a0932 · outbound
Which Viewpoint Shows it Best? Language for Weakly Supervising View Selection in Multi-view Instructional Videos Multi-view surveillance video summarization via joint embedding and sparse optimization
Reference 82
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 8f1a830f-752c-4357-a063-ac2da8b6188e · outbound
Which Viewpoint Shows it Best? Language for Weakly Supervising View Selection in Multi-view Instructional Videos Roy-Chowdhury
Reference 83
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation c21a3630-06c2-4b52-af5b-4584cb1ac0f1 · outbound
Which Viewpoint Shows it Best? Language for Weakly Supervising View Selection in Multi-view Instructional Videos Bleu: a method for automatic evaluation of machine translation
Reference 84
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 5a1a688b-0e61-4405-8da4-40835fb472d8 · outbound
Which Viewpoint Shows it Best? Language for Weakly Supervising View Selection in Multi-view Instructional Videos Sumgraph: Video summarization via recursive graph modeling
Reference 85
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 7d7ac8df-72f4-4c62-9db1-cc869eb1d79d · outbound
Which Viewpoint Shows it Best? Language for Weakly Supervising View Selection in Multi-view Instructional Videos Egovlpv2: Egocentric video-language pre-training with fusion in the backbone
Reference 86
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 378c0f34-b5b3-4334-b4f1-51e9efeac2e5 · outbound
Which Viewpoint Shows it Best? Language for Weakly Supervising View Selection in Multi-view Instructional Videos Vloc- net++: Deep multitask learning for semantic visual localiza- tion and odometry
Reference 87
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation a4a08e10-1781-4d09-a398-42aed2923331 · outbound
Which Viewpoint Shows it Best? Language for Weakly Supervising View Selection in Multi-view Instructional Videos Sidekick policy learning for active visual exploration
Reference 88
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation af605da0-9852-40f0-bac3-870705489a30 · outbound
Which Viewpoint Shows it Best? Language for Weakly Supervising View Selection in Multi-view Instructional Videos Emergence of exploratory look-around behaviors through active observation completion
Reference 89
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation e496eb70-24af-4f73-94f6-c2f539eb4e2a · outbound
Which Viewpoint Shows it Best? Language for Weakly Supervising View Selection in Multi-view Instructional Videos Naq: Leveraging narrations as queries to super- vise episodic memory
Reference 90
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation d19f8cfa-e3cc-4253-ad30-7ea6ae3fa489 · outbound
Which Viewpoint Shows it Best? Language for Weakly Supervising View Selection in Multi-view Instructional Videos Video summarization by learning from unpaired data
Reference 91
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 0d4725b1-c25f-457f-9870-5f756309db84 · outbound
Which Viewpoint Shows it Best? Language for Weakly Supervising View Selection in Multi-view Instructional Videos Adaptive video highlight detection by learning from user history
Reference 92
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2fc92f75-2740-4f13-a654-84a748667606 · outbound
Reference 93
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation c5a9690f-59df-487e-b0b1-bb2c99d7e9e3 · outbound
Which Viewpoint Shows it Best? Language for Weakly Supervising View Selection in Multi-view Instructional Videos Attend and segment: Attention guided active semantic segmentation
Reference 94
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 121c8f1a-836b-4abb-a24b-3acd089351ec · outbound
Which Viewpoint Shows it Best? Language for Weakly Supervising View Selection in Multi-view Instructional Videos Glimpse- attend-and-explore: Self-attention for active visual explo- ration
Reference 95
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation d6b9d89a-40b3-4049-8dfd-16e538200d9a · outbound
Which Viewpoint Shows it Best? Language for Weakly Supervising View Selection in Multi-view Instructional Videos Actor and observer: Joint modeling of first and third-person videos
Reference 96
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 8481d850-ddf3-403a-8c8c-56f49059cce9 · outbound
Which Viewpoint Shows it Best? Language for Weakly Supervising View Selection in Multi-view Instructional Videos Tvsum: Summarizing web videos using titles
Reference 97
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 6266f243-3dcb-4190-a5e2-0dda93bef23a · outbound
Which Viewpoint Shows it Best? Language for Weakly Supervising View Selection in Multi-view Instructional Videos Making 360 ° video watchable in 2d: Learning videography for click free view- ing
Reference 98
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 67a9f5c6-e0a6-41cf-9c32-e40ba2a166ae · outbound
Which Viewpoint Shows it Best? Language for Weakly Supervising View Selection in Multi-view Instructional Videos Pano2vid: Automatic cinematography for watching 360 videos
Reference 99
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation fbe32558-5b46-4a57-9a06-394c4ba8a462 · outbound
Which Viewpoint Shows it Best? Language for Weakly Supervising View Selection in Multi-view Instructional Videos Automatic con- cept discovery from parallel text and visual corpora
Reference 100
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 6cd707ae-c6d6-4b32-b3a1-eec071956336 · inbound
Switch-a-View: View Selection Learned from Unlabeled In-the-wild Videos Which Viewpoint Shows it Best? Language for Weakly Supervising View Selection in Multi-view Instructional Videos
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ebe94ece-1e15-49a0-bcc6-e24f2cfb1f29 · inbound
Bridging Perspectives: A Survey on Cross-view Collaborative Intelligence with Egocentric-Exocentric Vision Which Viewpoint Shows it Best? Language for Weakly Supervising View Selection in Multi-view Instructional Videos
Reference 187
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.