Pith. sign in

Paper Citation Record · LEDGER

Generating 6DoF Object Manipulation Trajectories from Action Description in Egocentric Vision

As of 10 August 2026, this Paper Citation Record lists 100 of 111 outbound references and 1 inbound Pith citation observation for arXiv:2506.03605.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2506.03605 v1

Coverage vector

measured 100 of 111 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T11:03:01.873084Z

measured 101 of 101 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-10T06:31:04.303077+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-01T21:08:03.385771Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-08-05T02:28:24.338817Z

Reference resolution

100 of 111 outbound references displayed

  • verified exact0
  • verified fuzzy45
  • unresolved55
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

0
pith, observed 2026-08-05T02:28:24.338817Z

Outbound references

Observation 47410969-3493-4621-8539-833a4921c801 · outbound

This paper cites GPT-4 Technical Report.

Generating 6DoF Object Manipulation Trajectories from Action Description in Egocentric Vision GPT-4 Technical Report

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-07T11:03:01.639168Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:03:01.639168Z digest=sha256:a2810921509722f553179ba9dd648b2e1d20f362ffcc5e358a78629b5e688c87

Observation 0e0448d2-0fd0-46e8-b11f-0be3ee0c163c · outbound

This paper cites Affordances from human videos as a versatile representation for robotics.

Generating 6DoF Object Manipulation Trajectories from Action Description in Egocentric Vision Affordances from human videos as a versatile representation for robotics

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-07T11:03:01.643039Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:03:01.643039Z digest=sha256:6a365351af5fe5801a50647230a8972179cc140335dadead8729ff06ad69e48b

Observation eacf354e-7547-4189-8c4f-a1aaa84966a9 · outbound

This paper cites Introducing HOT3D: An Egocentric Dataset for 3D Hand and Object Tracking.

Generating 6DoF Object Manipulation Trajectories from Action Description in Egocentric Vision Introducing HOT3D: An Egocentric Dataset for 3D Hand and Object Tracking

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-07T11:03:01.645783Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:03:01.645783Z digest=sha256:f731de41e58c33e9f5d96315bbfc635f175778e4d657ecbd7b220af136a2a7d7

Observation c5de6b31-54c3-43ed-8559-99d5c96e98dc · outbound

This paper cites METEOR: An auto- matic metric for MT evaluation with improved correlation with human judgments.

Generating 6DoF Object Manipulation Trajectories from Action Description in Egocentric Vision METEOR: An auto- matic metric for MT evaluation with improved correlation with human judgments

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-07T11:03:01.648655Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:03:01.648655Z digest=sha256:d4624a94077c637a65e452353e9457f4be5887e52beb7956575bb1321c680327

Observation 11cb9d9e-ebb4-4334-ba79-b778064223d4 · outbound

This paper cites Uncertainty-aware state space transformer for egocentric 3d hand trajectory forecasting.

Generating 6DoF Object Manipulation Trajectories from Action Description in Egocentric Vision Uncertainty-aware state space transformer for egocentric 3d hand trajectory forecasting

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-07T11:03:01.651358Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:03:01.651358Z digest=sha256:77b119559ffd9e6272bedb5537670a4775d9a375825005da9b2f7013db3b8624

Observation de485560-fb69-47ec-9292-dd7e502dd5d2 · outbound

This paper cites ARKitscenes - a diverse real-world dataset for 3d indoor scene understanding using mobile RGB-d data.

Generating 6DoF Object Manipulation Trajectories from Action Description in Egocentric Vision ARKitscenes - a diverse real-world dataset for 3d indoor scene understanding using mobile RGB-d data

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-07T11:03:01.653992Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:03:01.653992Z digest=sha256:f0cff1506ca7bb8469ee1cd3f68a2836337fa726b1916a0f3837fc34c6dc6802

Observation b1cea614-4c66-407d-9791-5c25fa705b26 · outbound

This paper cites Yu, and Jianbo Shi.

Generating 6DoF Object Manipulation Trajectories from Action Description in Egocentric Vision Yu, and Jianbo Shi

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-07T11:03:01.656749Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:03:01.656749Z digest=sha256:08fa4ed0edae9fc939b2ec5fdec2b489663fd2c5d8782d4328aa26d16938bf72

Observation 92a0aeb6-a92e-4280-802d-537d3320f81e · outbound

This paper cites Kemp, and James Hays.

Generating 6DoF Object Manipulation Trajectories from Action Description in Egocentric Vision Kemp, and James Hays

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-07T11:03:01.659454Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:03:01.659454Z digest=sha256:822b868f9a0805782f985c7ac9e2e838685efe7e25e31ca04b7dea3499f18657

Observation b2979133-5d9f-4006-9856-5fe7c1fc25cc · outbound

This paper cites RT-1: Robotics Transformer for Real- World Control at Scale.

Generating 6DoF Object Manipulation Trajectories from Action Description in Egocentric Vision RT-1: Robotics Transformer for Real- World Control at Scale

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-07T11:03:01.662279Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:03:01.662279Z digest=sha256:a2c993c105a64c483f3d7ceb6d2b1f4cd6df3a8f0c4e8f35ad09e9d6c66b6276

Observation 4c9c4bda-f8b5-4526-af9a-3b52d7b56fb1 · outbound

This paper cites Deep regression on manifolds: A 3d ro- tation case study.

Generating 6DoF Object Manipulation Trajectories from Action Description in Egocentric Vision Deep regression on manifolds: A 3d ro- tation case study

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-07T11:03:01.665148Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:03:01.665148Z digest=sha256:9ac259b8276f9b82f2dba976903cb1191986f0bcc74ae959b9666120e524078b

Observation 557b2e18-0eaf-43e4-ba88-bce57396ee90 · outbound

This paper cites Text2hoi: Text-guided 3d motion generation for hand-object interaction.

Generating 6DoF Object Manipulation Trajectories from Action Description in Egocentric Vision Text2hoi: Text-guided 3d motion generation for hand-object interaction

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-07T11:03:01.667377Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:03:01.667377Z digest=sha256:c9cbcb568206c43ae1644463309ce7dded33d84c79f3a6be44a72e7796b665de

Observation 561ecdfc-d092-4cf9-9d8b-f12debce7731 · outbound

This paper cites Fleet, and Geoffrey Hinton.

Generating 6DoF Object Manipulation Trajectories from Action Description in Egocentric Vision Fleet, and Geoffrey Hinton

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-07T11:03:01.669846Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:03:01.669846Z digest=sha256:afa4d4b73f98b2b69c6f39deb1e2cdb8acd74f571c300213752d5ed30da54b6b

Observation ea0f3456-6039-4168-b018-0686df5c26ad · outbound

This paper cites Looking to relations for future trajectory forecast.

Generating 6DoF Object Manipulation Trajectories from Action Description in Egocentric Vision Looking to relations for future trajectory forecast

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-07T11:03:01.672243Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:03:01.672243Z digest=sha256:eb4261ff56743f110e925fa2b57711a158e51e955d9f8584456a7cdbe3959b59

Observation 7e55b631-8c15-421e-bff0-bcadd4487b6d · outbound

This paper cites Learning to act properly: Predicting and explain- ing affordances from images.

Generating 6DoF Object Manipulation Trajectories from Action Description in Egocentric Vision Learning to act properly: Predicting and explain- ing affordances from images

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-07T11:03:01.674396Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:03:01.674396Z digest=sha256:1c84fdabd68013d07d40f39d22f5b2f375210df47579cc01cd19935ca11fc556

Observation a68d4040-8d7d-4549-97c3-31865a1c710a · outbound

This paper cites Open x-embodiment: Robotic learning datasets and rt-x models: Open x- embodiment collaboration.

Generating 6DoF Object Manipulation Trajectories from Action Description in Egocentric Vision Open x-embodiment: Robotic learning datasets and rt-x models: Open x- embodiment collaboration

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-07T11:03:01.676590Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:03:01.676590Z digest=sha256:d4562ecbcf5d92bec50e4228c412d66e8c1a78bc166b504ad68bfbaadea64f94

Observation 27c3dced-da60-421b-bc08-b56860c1b583 · outbound

This paper cites Ganhand: Predicting human grasp affordances in multi-object scenes.

Generating 6DoF Object Manipulation Trajectories from Action Description in Egocentric Vision Ganhand: Predicting human grasp affordances in multi-object scenes

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-07T11:03:01.678794Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:03:01.678794Z digest=sha256:b7c842873bff713b458812c9df73aec9f818bd211b4d20481426cee1bddc2825

Observation 1fb4cb0e-aa1b-4380-a678-8667f639a872 · outbound

This paper cites Scaling egocentric vision: The epic- kitchens dataset.

Generating 6DoF Object Manipulation Trajectories from Action Description in Egocentric Vision Scaling egocentric vision: The epic- kitchens dataset

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-07T11:03:01.681123Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:03:01.681123Z digest=sha256:3bfb0f10db14bbe0ffec2f88c614dd8994c65ae2f8925203608f6422aa74bd94

Observation 5b0e67e0-837d-4c2a-8850-75b5b89d204e · outbound

This paper cites Rescaling egocentric vision: Collection, pipeline and challenges for epic-kitchens-100.Interna- tional Journal of Computer Vision, 130(1):33–55, 2022.

Generating 6DoF Object Manipulation Trajectories from Action Description in Egocentric Vision Rescaling egocentric vision: Collection, pipeline and challenges for epic-kitchens-100.Interna- tional Journal of Computer Vision, 130(1):33–55, 2022

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-07T11:03:01.683615Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:03:01.683615Z digest=sha256:34dc689d757af78736baf5e00010bfc0d3dbaff1d6a6ba7eeca837ae2c1d802c

Observation de27e247-7e19-4ed7-b155-7edbf824ff29 · outbound

This paper cites Affor- dancenet: An end-to-end deep learning approach for object affordance detection.

Generating 6DoF Object Manipulation Trajectories from Action Description in Egocentric Vision Affor- dancenet: An end-to-end deep learning approach for object affordance detection

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-07T11:03:01.685866Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:03:01.685866Z digest=sha256:0c89503cf0dc1c70375b5eb34058c3b7e412dd4defc08bb943b90c30bf05aac5

Observation cb20bff5-da7a-44e5-bcec-6300eccf0542 · outbound

This paper cites an unresolved cited work.

Generating 6DoF Object Manipulation Trajectories from Action Description in Egocentric Vision Unresolved cited work

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-07T11:03:01.688348Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:03:01.688348Z digest=sha256:c0dbe2e849c948bf47f89c153973b897d2ff81ddb13c9161f0fdb8295f0b9b86

Observation f70df54e-48eb-4c63-b46c-d28fefc7c550 · outbound

This paper cites Project Aria: A New Tool for Egocentric Multi-Modal AI Research.

Generating 6DoF Object Manipulation Trajectories from Action Description in Egocentric Vision Project Aria: A New Tool for Egocentric Multi-Modal AI Research

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-07T11:03:01.690863Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:03:01.690863Z digest=sha256:7b52f98db1527511a99bb9913904a1de51a8f61c0b0e8bcd0b063c932e4586b7

Observation 4cafee9a-65f6-46df-be94-96f49ae2bf13 · outbound

This paper cites Egopat3dv2: Predicting 3d action target from 2d egocentric vision for human-robot interac- tion.

Generating 6DoF Object Manipulation Trajectories from Action Description in Egocentric Vision Egopat3dv2: Predicting 3d action target from 2d egocentric vision for human-robot interac- tion

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-07T11:03:01.693195Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:03:01.693195Z digest=sha256:bff04d7eb82064d1ec5fc2996eea16bf3274250f547b9cd32228f69c964939a4

Observation 4f40752f-fb0c-4820-8e37-cacb93de807b · outbound

This paper cites Fischler and Robert C.

Generating 6DoF Object Manipulation Trajectories from Action Description in Egocentric Vision Fischler and Robert C

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-07T11:03:01.695507Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:03:01.695507Z digest=sha256:a6b42f1322859a17589b838ae12ef38d76b2b7384ee945d982477e516e857905

Observation 822acb03-881c-4f2b-9c00-033718bdcf53 · outbound

This paper cites Zhao, and Chelsea Finn.

Generating 6DoF Object Manipulation Trajectories from Action Description in Egocentric Vision Zhao, and Chelsea Finn

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-07T11:03:01.697706Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:03:01.697706Z digest=sha256:c9955862db3401874173a0fc832da482c4b0c48ef048b9ab2c4ac79e40a13c09

Observation bf437845-6a3b-4c46-846b-5ddb1e278a48 · outbound

This paper cites What would you expect? anticipating egocentric actions with rolling-unrolling lstms and modality attention.

Generating 6DoF Object Manipulation Trajectories from Action Description in Egocentric Vision What would you expect? anticipating egocentric actions with rolling-unrolling lstms and modality attention

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-07T11:03:01.699904Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:03:01.699904Z digest=sha256:6646381b722ec00403441c809d1f550a16e53dbd145699d02b141b03684124e3

Observation b43f8726-c20e-4eb5-b077-dd31dad3d841 · outbound

This paper cites Next-active-object predic- tion from egocentric videos.Journal of Visual Communi- cation and Image Representation, 49:401–411, 2017.

Generating 6DoF Object Manipulation Trajectories from Action Description in Egocentric Vision Next-active-object predic- tion from egocentric videos.Journal of Visual Communi- cation and Image Representation, 49:401–411, 2017

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-07T11:03:01.702296Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:03:01.702296Z digest=sha256:fa7a2fbd40b0b329ebe5196a67c1d0962b01a79022dc9a0ecd8fa9a3fe395545

Observation c23d645a-23b6-4211-ae47-869660f44c20 · outbound

This paper cites So predictable! continuous 3d hand trajectory prediction in virtual reality.

Generating 6DoF Object Manipulation Trajectories from Action Description in Egocentric Vision So predictable! continuous 3d hand trajectory prediction in virtual reality

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-07T11:03:01.705143Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:03:01.705143Z digest=sha256:64088adf119c390fe474986492db96aeef24bed547aa05dd41d2e3e0c1ff69ec

Observation df8fd2ba-d4cd-4342-a041-96d9b3477dc5 · outbound

This paper cites Transformer networks for trajectory forecasting.

Generating 6DoF Object Manipulation Trajectories from Action Description in Egocentric Vision Transformer networks for trajectory forecasting

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-07T11:03:01.707448Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:03:01.707448Z digest=sha256:6c500522f2e2335685dceb8a49189895b39cbff3c2eac4ea57a1290fe25f3bc8

Observation 38f75670-71f1-4db6-bf5a-88f0d6751f35 · outbound

This paper cites Ego4d: Around the world in 3,000 hours of egocentric video.

Generating 6DoF Object Manipulation Trajectories from Action Description in Egocentric Vision Ego4d: Around the world in 3,000 hours of egocentric video

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-07T11:03:01.709696Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:03:01.709696Z digest=sha256:27797e68ec651a319ffd63e943e868663a74786f43d8c19f1c5222ba20fb65db

Observation 8af6e787-429b-47bc-981c-a3128c1425d3 · outbound

This paper cites Ego-Exo4D: Understanding Skilled Human Activity from First- and Third-Person Perspectives.

Generating 6DoF Object Manipulation Trajectories from Action Description in Egocentric Vision Ego-Exo4D: Understanding Skilled Human Activity from First- and Third-Person Perspectives

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-07T11:03:01.711729Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:03:01.711729Z digest=sha256:d151c77e1ba62f8939db13dc0677e8104ea69f077d88dc8edcf66a0a3de234ae

Observation 21d1c34b-ed98-4cfd-8483-993549dd4e4a · outbound

This paper cites Handal: A dataset of real-world manipulable object cate- gories with pose annotations, affordances, and reconstruc- tions.

Generating 6DoF Object Manipulation Trajectories from Action Description in Egocentric Vision Handal: A dataset of real-world manipulable object cate- gories with pose annotations, affordances, and reconstruc- tions

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-07T11:03:01.714360Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:03:01.714360Z digest=sha256:85c3fd480671692cac5a700aee572b0bc0297ac46e7edc3165485da7fec26d86

Observation 931b5475-e2fb-436a-a2d1-d96678ec075c · outbound

This paper cites Honnotate: A method for 3d annotation of hand and object poses.

Generating 6DoF Object Manipulation Trajectories from Action Description in Egocentric Vision Honnotate: A method for 3d annotation of hand and object poses

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-07T11:03:01.716689Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:03:01.716689Z digest=sha256:1478e07ebd9307af2d0a42cfb6f72a39e12a0bf3dfeeeae87c29d713977f4b88

Observation cb6f7cc7-ddf8-4477-837d-b16ebfd6c1fe · outbound

This paper cites Ego3dt: Tracking every 3d object in ego-centric videos.

Generating 6DoF Object Manipulation Trajectories from Action Description in Egocentric Vision Ego3dt: Tracking every 3d object in ego-centric videos

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-07T11:03:01.718900Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:03:01.718900Z digest=sha256:5c886c3b135624f0ff8b6a708575dfaa5ba30786b46a1984cb0d0fbeab2c7360

Observation adc330e8-6101-4ca2-9e30-b0f7e09ed2a3 · outbound

This paper cites Onepose++: Keypoint-free one- shot object pose estimation without cad models.

Generating 6DoF Object Manipulation Trajectories from Action Description in Egocentric Vision Onepose++: Keypoint-free one- shot object pose estimation without cad models

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-07T11:03:01.721146Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:03:01.721146Z digest=sha256:9a526d18bd719bba1d627c50a58de3229d240c65d17572004773d93abc194f08

Observation b4bc772e-e112-43ec-8d45-b39e3fa6dde2 · outbound

This paper cites Pvn3d: A deep point-wise 3d keypoints voting network for 6dof pose estimation.

Generating 6DoF Object Manipulation Trajectories from Action Description in Egocentric Vision Pvn3d: A deep point-wise 3d keypoints voting network for 6dof pose estimation

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-07T11:03:01.723553Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:03:01.723553Z digest=sha256:7bb2aea4da01963982fba628e807468676de09fffba655561c794dd2c9253095

Observation d14f3dfe-27d0-4714-aa2c-d098d497726d · outbound

This paper cites Fs6d: Few-shot 6d pose estimation of novel objects.

Generating 6DoF Object Manipulation Trajectories from Action Description in Egocentric Vision Fs6d: Few-shot 6d pose estimation of novel objects

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:03:02.834333Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T11:03:01.725649Z digest=sha256:834e166c4ff1d6bf8a894be971d8f38f4fdb580ff53a574c94c90e837bd15274

Observation fe9080e1-5070-4185-856b-56a5358a222d · outbound

This paper cites The curious case of neural text degeneration.

Generating 6DoF Object Manipulation Trajectories from Action Description in Egocentric Vision The curious case of neural text degeneration

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:03:02.827197Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T11:03:01.727761Z digest=sha256:c88f47e3ff9891eeea425780aaf6a3966bc9aa18e84e237c832d1b65e4d50f3f

Observation 17033dfd-1bed-4186-9d77-dc94c135f42a · outbound

This paper cites 3d-llm: Inject- ing the 3d world into large language models.

Generating 6DoF Object Manipulation Trajectories from Action Description in Egocentric Vision 3d-llm: Inject- ing the 3d world into large language models

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:03:02.820454Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T11:03:01.729959Z digest=sha256:ab17c3e362e88ad8db5f1921b6b07751fdc155e925d858198f4e33466eda8a14

Observation 04037fe9-cbc4-4c92-917d-e2c1e3160f1b · outbound

This paper cites LITA: Language Instructed Temporal-Localization Assistant.

Generating 6DoF Object Manipulation Trajectories from Action Description in Egocentric Vision LITA: Language Instructed Temporal-Localization Assistant

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-07T11:03:01.732235Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:03:01.732235Z digest=sha256:33af200937679d907c549a22d0179424e5e7c3055e8b4a11594afba42ad17c41

Observation 764af670-8aab-4d8c-bf47-021987343d08 · outbound

This paper cites V oxposer: Composable 3d value maps for robotic manipulation with language models.

Generating 6DoF Object Manipulation Trajectories from Action Description in Egocentric Vision V oxposer: Composable 3d value maps for robotic manipulation with language models

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:03:02.813619Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T11:03:01.734633Z digest=sha256:338b7c8264484c607e2a9a33fc90c0f859857f9a3c49b25357889279c5ffe3c3

Observation 2ef80ba7-fecf-4573-96ec-35d4f8b57f4e · outbound

This paper cites Technical Report for Ego4D Long Term Action Anticipation Challenge 2023.

Generating 6DoF Object Manipulation Trajectories from Action Description in Egocentric Vision Technical Report for Ego4D Long Term Action Anticipation Challenge 2023

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-07T11:03:01.737020Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:03:01.737020Z digest=sha256:68726749f420c3c306b06936c9cc90ba2090b0fc619551df2300554410c7c416

Observation f30ae56d-c238-4ccf-bb09-e808ef756a06 · outbound

This paper cites Jacobs, Michael I.

Generating 6DoF Object Manipulation Trajectories from Action Description in Egocentric Vision Jacobs, Michael I

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:03:02.806848Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T11:03:01.739593Z digest=sha256:fb4b186d0a6c8450408a1b91f35bc1320c1ff2bc750b722961b3d1a63f716429

Observation edbbc3bc-09a9-45cd-8a51-796d0fa0dfab · outbound

This paper cites Jordan and R.A.

Generating 6DoF Object Manipulation Trajectories from Action Description in Egocentric Vision Jordan and R.A

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:03:02.800121Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T11:03:01.741741Z digest=sha256:56276733fb47c4fbd30cc753df3f9cf4f36fb7526a389f4f948fa301362b17af

Observation 6771cede-c617-48fe-9099-84ac18712688 · outbound

This paper cites OpenVLA: An Open-Source Vision-Language-Action Model.

Generating 6DoF Object Manipulation Trajectories from Action Description in Egocentric Vision OpenVLA: An Open-Source Vision-Language-Action Model

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-07T11:03:01.743937Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:03:01.743937Z digest=sha256:c2b359d04e76d076118fb21996d70b3903742d41009c694f873c0189118af676

Observation c03a6023-fbd9-4c1a-b77a-4f4f58c3a29a · outbound

This paper cites Koppula and Ashutosh Saxena.

Generating 6DoF Object Manipulation Trajectories from Action Description in Egocentric Vision Koppula and Ashutosh Saxena

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:03:02.792891Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T11:03:01.746343Z digest=sha256:d5f102122b1ac32117c7e43b0134430f939a2e802256e56646fd1e7c108f6f78

Observation 7746be9a-727f-4292-9c42-34e4e355dd73 · outbound

This paper cites H2o: Two hands manipulating objects for first person interaction recognition.

Generating 6DoF Object Manipulation Trajectories from Action Description in Egocentric Vision H2o: Two hands manipulating objects for first person interaction recognition

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:03:02.785551Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T11:03:01.748950Z digest=sha256:3464180de026fb73410bb854f44fce84bc5b1aa6fec573d0a517a54f041928ec

Observation fd9ce56b-5902-45a9-9686-86d7318c9ae2 · outbound

This paper cites Choy, Philip H.

Generating 6DoF Object Manipulation Trajectories from Action Description in Egocentric Vision Choy, Philip H

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:03:02.778345Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T11:03:01.751582Z digest=sha256:317a7245196d773cbb3e65386c1385425d504c2c83f5ccca10808794d52fdf42

Observation 5d5f9080-3912-49cf-9ac3-2e07b2aebf5d · outbound

This paper cites Locate: Localize and transfer object parts for weakly supervised affordance grounding.

Generating 6DoF Object Manipulation Trajectories from Action Description in Egocentric Vision Locate: Localize and transfer object parts for weakly supervised affordance grounding

Reference 48

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:03:02.771210Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T11:03:01.753872Z digest=sha256:f5885b37e0259f93990508c205acb5383ca602b6ff5580299fde214977f810f1

Observation 6891d889-8271-47ca-9ab6-a7ea2e73b333 · outbound

This paper cites Learning precise affordances from ego- centric videos for robotic manipulation.arxiv preprint arXiv:2408.10123, 2024.

Generating 6DoF Object Manipulation Trajectories from Action Description in Egocentric Vision Learning precise affordances from ego- centric videos for robotic manipulation.arxiv preprint arXiv:2408.10123, 2024

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-07T11:03:01.755952Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:03:01.755952Z digest=sha256:424db7578986845d334164362e0787f882f4e751511347cf46722ad3bb4bedb4

Observation 2a635c79-6523-4826-af66-c6660f6a560c · outbound

This paper cites BLIP-2: Bootstrapping language-image pre-training with frozen image encoders and large language models.

Generating 6DoF Object Manipulation Trajectories from Action Description in Egocentric Vision BLIP-2: Bootstrapping language-image pre-training with frozen image encoders and large language models

Reference 50

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:03:02.763851Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T11:03:01.758183Z digest=sha256:21fba08c171e351c2b56fe44e52a25e4b76af651f3750b59a5f776dfca469151

Observation ccb09b42-4bb9-4ce6-abcb-4a4662bc6d85 · outbound

This paper cites an unresolved cited work.

Generating 6DoF Object Manipulation Trajectories from Action Description in Egocentric Vision Unresolved cited work

Reference 51

Resolution
unresolved
raw_fallback, observed 2026-08-07T11:03:02.756194Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T11:03:01.760603Z digest=sha256:55b299da5d2a7214c721c32fa7255af9c281c41548905d579a7e94b14c2240a4

Observation 3837d5c2-7a9c-4e69-941e-4ace089e592c · outbound

This paper cites Deepim: Deep iterative matching for 6d pose estimation.

Generating 6DoF Object Manipulation Trajectories from Action Description in Egocentric Vision Deepim: Deep iterative matching for 6d pose estimation

Reference 52

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:03:02.749185Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T11:03:01.762807Z digest=sha256:5750cc0cc21cc31da887e99702a3637ca56a85b07c3ea310574325788c325caf

Observation e991d2d3-bc65-4b85-998d-397973cb58d8 · outbound

This paper cites Egocentric predic- tion of action target in 3d.

Generating 6DoF Object Manipulation Trajectories from Action Description in Egocentric Vision Egocentric predic- tion of action target in 3d

Reference 53

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:03:02.741266Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T11:03:01.765023Z digest=sha256:b62cff6ae9bc3fc07ca5a48a9466068462422db3dbf2a5ddc4639fc25c181205

Observation 4b58a739-6ba0-40d7-b793-ad66de36a381 · outbound

This paper cites Vila: On pre-training for vi- sual language models.

Generating 6DoF Object Manipulation Trajectories from Action Description in Egocentric Vision Vila: On pre-training for vi- sual language models

Reference 54

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:03:02.733369Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T11:03:01.767024Z digest=sha256:3b3fb03238051933a5b52b51c5354b4cf18e7365b5a46272f5753ba196d4645b

Observation 1aaac6ff-ea91-4329-a438-0c2fedd866c8 · outbound

This paper cites Visual instruction tuning.

Generating 6DoF Object Manipulation Trajectories from Action Description in Egocentric Vision Visual instruction tuning

Reference 55

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:03:02.725673Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T11:03:01.769093Z digest=sha256:64a4d389e42b9f6c8c1801f149577614a26f52cf3171574153b0c7843bfe86b2

Observation c9bba91b-93f4-4abe-8075-a3ad36481ef3 · outbound

This paper cites Improved baselines with visual instruction tuning.

Generating 6DoF Object Manipulation Trajectories from Action Description in Egocentric Vision Improved baselines with visual instruction tuning

Reference 56

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:03:02.718489Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T11:03:01.771408Z digest=sha256:9a243befcdae02c0f3f9e9b4829583c062f0b1662d6e57dff4d139d4dc8b812f

Observation 95b62291-4839-402c-95b4-e81f6f958610 · outbound

This paper cites an unresolved cited work.

Generating 6DoF Object Manipulation Trajectories from Action Description in Egocentric Vision Unresolved cited work

Reference 57

Resolution
unresolved
raw_fallback, observed 2026-08-07T11:03:02.711058Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T11:03:01.773689Z digest=sha256:fb1cdd79db581680144f7131a686a841fc0579e65c6d241ab3ef3f720db0770c

Observation 89577c0a-d18c-4960-8da4-d952bdd5d7ef · outbound

This paper cites Joint hand motion and interaction hotspots prediction from egocentric videos.

Generating 6DoF Object Manipulation Trajectories from Action Description in Egocentric Vision Joint hand motion and interaction hotspots prediction from egocentric videos

Reference 58

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:03:02.703822Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T11:03:01.776268Z digest=sha256:8228c743c635992c5b6b56a5357623b471e6d7370e6693aa9cff9dd11436b6ab

Observation 447f20c7-76fc-47ae-9800-9c979c79da83 · outbound

This paper cites Grounding DINO: Marrying DINO with Grounded Pre-Training for Open-Set Object Detection.

Generating 6DoF Object Manipulation Trajectories from Action Description in Egocentric Vision Grounding DINO: Marrying DINO with Grounded Pre-Training for Open-Set Object Detection

Reference 59

Resolution
unresolved
no resolver link, observed 2026-08-07T11:03:01.778564Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:03:01.778564Z digest=sha256:f62fc952d3a744420e13145f171b5703d6738fd22af052b5352673a2b25598b4

Observation 023e3407-0794-4240-b2a2-df951c391638 · outbound

This paper cites Hoi4d: A 4d egocentric dataset for category-level human-object interaction.

Generating 6DoF Object Manipulation Trajectories from Action Description in Egocentric Vision Hoi4d: A 4d egocentric dataset for category-level human-object interaction

Reference 60

Resolution
unresolved
no resolver link, observed 2026-08-07T11:03:01.781185Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:03:01.781185Z digest=sha256:07ad71076febae8bac7aca43dca9d3a0225cb3e369efe28107db7e5a825c3601

Observation e08fbb80-4ef6-4756-93f0-7476aee864c0 · outbound

This paper cites Decoupled weight decay regularization.

Generating 6DoF Object Manipulation Trajectories from Action Description in Egocentric Vision Decoupled weight decay regularization

Reference 61

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:03:02.691148Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T11:03:01.783372Z digest=sha256:1e1e87a8d94d859dadba0fe485fc94acd71e138578d0871afecdd7ffd5271752

Observation 930debae-048d-41d0-a0eb-29c7346224a7 · outbound

This paper cites Phrase-based affordance detection via cyclic bi- lateral interaction.IEEE Transactions on Artificial Intelli- gence, 4(5):1186–1198, 2023.

Generating 6DoF Object Manipulation Trajectories from Action Description in Egocentric Vision Phrase-based affordance detection via cyclic bi- lateral interaction.IEEE Transactions on Artificial Intelli- gence, 4(5):1186–1198, 2023

Reference 62

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:03:02.684074Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T11:03:01.785682Z digest=sha256:fb1d24fb8ac79be7f65b0b114c400ff6be9a5ca52e638d40ea68acc1dd150b5c

Observation f15860b1-d882-4991-b12c-21228494844b · outbound

This paper cites Madiff: Motion-aware mamba diffusion models for hand trajectory prediction on egocentric videos.

Generating 6DoF Object Manipulation Trajectories from Action Description in Egocentric Vision Madiff: Motion-aware mamba diffusion models for hand trajectory prediction on egocentric videos

Reference 63

Resolution
unresolved
no resolver link, observed 2026-08-07T11:03:01.787848Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:03:01.787848Z digest=sha256:360fd253b9621af122220165ae12d6f02901ac009c68347cff46c39546380229

Observation 2026da06-56bc-47f1-846b-17d4caab190a · outbound

This paper cites Diff-ip2d: Diffusion-based hand-object interac- tion prediction on egocentric videos.arXiv preprint arXiv:2405.04370, 2024.

Generating 6DoF Object Manipulation Trajectories from Action Description in Egocentric Vision Diff-ip2d: Diffusion-based hand-object interac- tion prediction on egocentric videos.arXiv preprint arXiv:2405.04370, 2024

Reference 64

Resolution
unresolved
no resolver link, observed 2026-08-07T11:03:01.789958Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:03:01.789958Z digest=sha256:01086e3f36ecfe2405157ad460c13b2034c5ef2ba1862a7b15a425c7d25fdedb

Observation 5153fdae-f7e7-4702-99b5-59a4e0fca411 · outbound

This paper cites Quest 3, 2023.

Generating 6DoF Object Manipulation Trajectories from Action Description in Egocentric Vision Quest 3, 2023

Reference 65

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:03:02.676435Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T11:03:01.792768Z digest=sha256:e04ed4cefb510e47f67638b3f9fde34ca531cb8dfbc50f134a1996dc4b142d8b

Observation 819c8625-4a37-4003-b1c7-e15ea6411a84 · outbound

This paper cites Leveraging the present to anticipate the future in videos.

Generating 6DoF Object Manipulation Trajectories from Action Description in Egocentric Vision Leveraging the present to anticipate the future in videos

Reference 66

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:03:02.668973Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T11:03:01.795177Z digest=sha256:e5328c951e689d5bca201684b6935bf28ef284e9e275e1fbc93f9334f85be8ab

Observation 1102a049-928a-4b00-8f73-b05afd3ba384 · outbound

This paper cites RT-Affordance: Affordances are Versatile Intermediate Representations for Robot Manipulation.

Generating 6DoF Object Manipulation Trajectories from Action Description in Egocentric Vision RT-Affordance: Affordances are Versatile Intermediate Representations for Robot Manipulation

Reference 67

Resolution
unresolved
no resolver link, observed 2026-08-07T11:03:01.797316Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:03:01.797316Z digest=sha256:19796b55e646507186e2c482e071726b7e26d0b589ff289cf820f4e46a448aff

Observation d6df2c9a-dc7a-42bc-8c84-71a04c76e048 · outbound

This paper cites Open-vocabulary affordance detection in 3d point clouds.

Generating 6DoF Object Manipulation Trajectories from Action Description in Egocentric Vision Open-vocabulary affordance detection in 3d point clouds

Reference 68

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:03:02.593310Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T11:03:01.799828Z digest=sha256:0300c34af61aac6bec18763ad59a0dbb8be75f93744f7031ae9861fa9b60055d

Observation 8c71b377-7abe-4271-8ff7-7e4f1e62d16e · outbound

This paper cites Bleu: a method for automatic evaluation of machine translation.

Generating 6DoF Object Manipulation Trajectories from Action Description in Egocentric Vision Bleu: a method for automatic evaluation of machine translation

Reference 69

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:03:02.585613Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T11:03:01.802108Z digest=sha256:75da318e4bb9e3a5cedce74a44dd2be40716b9de27664eab31f7d32876360178

Observation 42a10d2e-0ed0-4b9e-a118-10202a618bf0 · outbound

This paper cites Colored point cloud registration revisited.

Generating 6DoF Object Manipulation Trajectories from Action Description in Egocentric Vision Colored point cloud registration revisited

Reference 70

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:03:02.578279Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T11:03:01.804235Z digest=sha256:a0d212ce073f49a13dc4a2e9d1eef6e7633422b3ddb6a26950e9f3e7107a6738

Observation 7b1dbff0-fffd-40ef-92ba-84967b0d8a59 · outbound

This paper cites Pix2pose: Pixel-wise coordinate regression of objects for 6d pose es- timation.

Generating 6DoF Object Manipulation Trajectories from Action Description in Egocentric Vision Pix2pose: Pixel-wise coordinate regression of objects for 6d pose es- timation

Reference 71

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:03:02.570682Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T11:03:01.806379Z digest=sha256:d06f021fb69fd219fe3a5718a1770f64c4822ba5a42473c0610c5c24af9b3415

Observation 8aab98a6-a389-4042-a712-40a55c6136f9 · outbound

This paper cites GloVe: Global vectors for word representation.

Generating 6DoF Object Manipulation Trajectories from Action Description in Egocentric Vision GloVe: Global vectors for word representation

Reference 72

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:03:02.561159Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T11:03:01.808650Z digest=sha256:4633dfd8d1ac72dd6777ddd4c09e8541656b041b6f59abb00d64a83c374766d8

Observation 27344cf1-4e69-451c-be86-c28f395af6e0 · outbound

This paper cites Detecting activities of daily living in first-person camera views.

Generating 6DoF Object Manipulation Trajectories from Action Description in Egocentric Vision Detecting activities of daily living in first-person camera views

Reference 73

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:03:02.553759Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T11:03:01.810884Z digest=sha256:e1dba986ca532d345680c36556cfd750dc06a395f0cfaaa65d6fef898b5c4971

Observation 5ea1d6b9-698f-42d3-8f70-1c24eae0f2bc · outbound

This paper cites Spatial Cognition from Egocentric Video: Out of Sight, Not Out of Mind.

Generating 6DoF Object Manipulation Trajectories from Action Description in Egocentric Vision Spatial Cognition from Egocentric Video: Out of Sight, Not Out of Mind

Reference 74

Resolution
unresolved
no resolver link, observed 2026-08-07T11:03:01.813217Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:03:01.813217Z digest=sha256:2d3351c4520d37be253a4ffd0df50fc7a4eb36732929ae3c1bdb096cd595b555

Observation fa4af39c-dada-4701-bfa9-78b139a080aa · outbound

This paper cites Learning transferable visual models from natural language supervision.

Generating 6DoF Object Manipulation Trajectories from Action Description in Egocentric Vision Learning transferable visual models from natural language supervision

Reference 75

Resolution
unresolved
no resolver link, observed 2026-08-07T11:03:01.815569Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:03:01.815569Z digest=sha256:8801064722e0ead42c74c88f27a9671822c873c8efc7493db4eedc42ea445a3d

Observation ed255f1f-fd84-435a-8580-9eb8727f6062 · outbound

This paper cites Grounded SAM: Assembling Open-World Models for Diverse Visual Tasks.

Generating 6DoF Object Manipulation Trajectories from Action Description in Egocentric Vision Grounded SAM: Assembling Open-World Models for Diverse Visual Tasks

Reference 76

Resolution
unresolved
no resolver link, observed 2026-08-07T11:03:01.817785Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:03:01.817785Z digest=sha256:f02eefc81faab67cb2395a4174d47ac80803d0a1db5b661a730c2a147f578311

Observation 81922796-4871-435b-9a34-03e8d7c49c91 · outbound

This paper cites Action scene graphs for long-form understanding of egocentric videos.

Generating 6DoF Object Manipulation Trajectories from Action Description in Egocentric Vision Action scene graphs for long-form understanding of egocentric videos

Reference 77

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:03:02.541814Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T11:03:01.820214Z digest=sha256:b36d8e9f386b2dbd36c116f408a22b66ee44041edc93b194e95705b5ac9fb2bc

Observation 14504436-79ae-42d3-9260-94546030c3fc · outbound

This paper cites Fast point feature histograms (fpfh) for 3d registration.

Generating 6DoF Object Manipulation Trajectories from Action Description in Egocentric Vision Fast point feature histograms (fpfh) for 3d registration

Reference 78

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:03:02.534947Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T11:03:01.823089Z digest=sha256:0ba202f02748a4b24db3f0b81523ca1df9a68769877d579ee1654e4ef80a1abe

Observation ade2407b-596b-4b88-8b2d-c95a416708f2 · outbound

This paper cites What object should i use? - task driven object detection.

Generating 6DoF Object Manipulation Trajectories from Action Description in Egocentric Vision What object should i use? - task driven object detection

Reference 79

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:03:02.527767Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T11:03:01.825425Z digest=sha256:43c9799d4318a38682d2002bb393f5cab17692235ad01649b6ae0dd39cef18ca

Observation 1ea84205-78c7-42aa-a7b9-ebb01551aaca · outbound

This paper cites As- sembly101: A large-scale multi-view video dataset for un- derstanding procedural activities.

Generating 6DoF Object Manipulation Trajectories from Action Description in Egocentric Vision As- sembly101: A large-scale multi-view video dataset for un- derstanding procedural activities

Reference 80

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:03:02.520789Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T11:03:01.827561Z digest=sha256:5b1283ac9d1d7062cda803bc6991d03ab5a5e8297c6ef3f7675509d9628549fe

Observation 02e18a5e-1626-432f-9be8-acb4e358afcc · outbound

This paper cites an unresolved cited work.

Generating 6DoF Object Manipulation Trajectories from Action Description in Egocentric Vision Unresolved cited work

Reference 81

Resolution
unresolved
raw_fallback, observed 2026-08-07T11:03:02.513616Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T11:03:01.829880Z digest=sha256:291381dc52c08a9c22c23325eb4c7aa75db8e7df9556e3f0c6b1a7ae6f1d6d70

Observation 03be063f-c5ca-4042-9db8-e65b21cf0d39 · outbound

This paper cites Ego4d goal-step: Toward hierarchical understanding of procedural activities.

Generating 6DoF Object Manipulation Trajectories from Action Description in Egocentric Vision Ego4d goal-step: Toward hierarchical understanding of procedural activities

Reference 82

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:03:02.506891Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T11:03:01.832360Z digest=sha256:f11d47bd143a7f99747e8ece0ba1daa924bd3caefbb0aa88055e68a26db5416a

Observation d5f21668-f04e-48dd-a9d6-ae5ca92bef9b · outbound

This paper cites Onepose: One-shot object pose estimation without cad models.

Generating 6DoF Object Manipulation Trajectories from Action Description in Egocentric Vision Onepose: One-shot object pose estimation without cad models

Reference 83

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:03:02.499917Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T11:03:01.834662Z digest=sha256:fef6a247ef4c54dc4d7aef29694e41f51052ad6053b7a6b8f9a1d0004036b0aa

Observation fba8b8e1-790e-440e-895c-388bb1383650 · outbound

This paper cites an unresolved cited work.

Generating 6DoF Object Manipulation Trajectories from Action Description in Egocentric Vision Unresolved cited work

Reference 84

Resolution
unresolved
raw_fallback, observed 2026-08-07T11:03:02.493194Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T11:03:01.836944Z digest=sha256:b7c442fd9f72f8f856996635ce3431c780c23ba04eb0df24ee631e5d0a0e80fd

Observation 57486b92-00fb-424a-868f-721d11782568 · outbound

This paper cites MiniGPT-3D: Efficiently Aligning 3D Point Clouds with Large Language Models using 2D Priors.

Generating 6DoF Object Manipulation Trajectories from Action Description in Egocentric Vision MiniGPT-3D: Efficiently Aligning 3D Point Clouds with Large Language Models using 2D Priors

Reference 85

Resolution
unresolved
no resolver link, observed 2026-08-07T11:03:01.839325Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:03:01.839325Z digest=sha256:9e428a6a25264d41a51ba820d848572d2a212acc40a9bebe6bbc6d296cec8bdd

Observation d8c0f507-059b-4f64-a679-6558183364cf · outbound

This paper cites Leveraging next-active ob- jects for context-aware anticipation in egocentric videos.

Generating 6DoF Object Manipulation Trajectories from Action Description in Egocentric Vision Leveraging next-active ob- jects for context-aware anticipation in egocentric videos

Reference 86

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:03:02.486631Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T11:03:01.841611Z digest=sha256:a3902393aae66fc7fb38ae619a484b36224e6f059b0ff47034d8647be509e2d0

Observation ad808590-b6e7-45c3-bb81-048eb365dfbf · outbound

This paper cites LLaMA: Open and Efficient Foundation Language Models.

Generating 6DoF Object Manipulation Trajectories from Action Description in Egocentric Vision LLaMA: Open and Efficient Foundation Language Models

Reference 87

Resolution
unresolved
no resolver link, observed 2026-08-07T11:03:01.843780Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:03:01.843780Z digest=sha256:ce34e09439d6206e2cf42dcccfeac1471cd0c28ecbcafcd2ed3b5922ed4edfb1

Observation 4e70f3ce-e83f-4137-9b97-2d60b237b59d · outbound

This paper cites Epic fields: Marrying 3d geometry and video un- derstanding.

Generating 6DoF Object Manipulation Trajectories from Action Description in Egocentric Vision Epic fields: Marrying 3d geometry and video un- derstanding

Reference 88

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:03:02.479483Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T11:03:01.845967Z digest=sha256:e828c718d1b77b1b034472c15428239447ac04d73adb78b455333fe2819fa6a6

Observation dfc00f89-07c2-4bea-8149-091d1f37e093 · outbound

This paper cites Attention is all you need.

Generating 6DoF Object Manipulation Trajectories from Action Description in Egocentric Vision Attention is all you need

Reference 89

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:03:02.472614Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T11:03:01.848178Z digest=sha256:0b9d432278e4ab4a4978b2ce81dad93e8f2e22a48b7957fb3e2e0d41abf3916f

Observation 765a3edd-8ca3-45de-94f2-35cdd54d589b · outbound

This paper cites Omnivid: A gener- ative framework for universal video understanding.

Generating 6DoF Object Manipulation Trajectories from Action Description in Egocentric Vision Omnivid: A gener- ative framework for universal video understanding

Reference 90

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:03:02.465495Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T11:03:01.850456Z digest=sha256:2c0a691f780a4fede93e9a2dfe8be1d6f2769acae02297bb6088b4ef6c091f40

Observation e873cb05-ef21-4921-9033-1513b53c6e49 · outbound

This paper cites OFA: Unifying architectures, tasks, and modalities through a simple sequence-to-sequence learning framework.

Generating 6DoF Object Manipulation Trajectories from Action Description in Egocentric Vision OFA: Unifying architectures, tasks, and modalities through a simple sequence-to-sequence learning framework

Reference 91

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:03:02.458516Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T11:03:01.852711Z digest=sha256:5cc3e1e344f94578df011568fd5b361eeef6ce1229e917fb9449dcecc5263f09

Observation ed3fb1cf-8abe-4e6b-951d-0cf6523a988b · outbound

This paper cites Dust3r: Geometric 3d vision made easy.

Generating 6DoF Object Manipulation Trajectories from Action Description in Egocentric Vision Dust3r: Geometric 3d vision made easy

Reference 92

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:03:02.451591Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T11:03:01.854962Z digest=sha256:95bc06ca45e261a450fbd90ef99e454625a1b0118c5d94f14016b0a0e5746e09

Observation ad6ba276-8925-4a1c-820e-bd952f9dd376 · outbound

This paper cites Holoassist: an egocentric human interaction dataset for interactive ai assistants in the real world.

Generating 6DoF Object Manipulation Trajectories from Action Description in Egocentric Vision Holoassist: an egocentric human interaction dataset for interactive ai assistants in the real world

Reference 93

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:03:02.444616Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T11:03:01.857232Z digest=sha256:a54989bd2ae0071bd9e04447632dad19af8d5ecf7ba7c96d4f0c3eb2b57d98c0

Observation f21af856-db0f-4b4f-8b68-447966d98669 · outbound

This paper cites Foundationpose: Unified 6d pose estimation and tracking of novel objects.

Generating 6DoF Object Manipulation Trajectories from Action Description in Egocentric Vision Foundationpose: Unified 6d pose estimation and tracking of novel objects

Reference 94

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:03:02.437955Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T11:03:01.859550Z digest=sha256:48ffc1938be79b7551e3d3430cf0e8df58b23b7c251f7e9b138d4cf114c1691f

Observation d361858b-5ecf-4bad-b00d-82f773432897 · outbound

This paper cites Learning descriptors for object recognition and 3d pose estimation.

Generating 6DoF Object Manipulation Trajectories from Action Description in Egocentric Vision Learning descriptors for object recognition and 3d pose estimation

Reference 95

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:03:02.430859Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T11:03:01.861969Z digest=sha256:efaf5c008597919285bf1d24c5d042dcf09618b56ea1a0cd98120cfdacf814d3

Observation 9ec32dc1-f846-42df-8dda-4e1cd383b105 · outbound

This paper cites Spatialtracker: Tracking any 2d pixels in 3d space.

Generating 6DoF Object Manipulation Trajectories from Action Description in Egocentric Vision Spatialtracker: Tracking any 2d pixels in 3d space

Reference 96

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:03:02.423930Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T11:03:01.864022Z digest=sha256:ea11b18ec081b7ecb7dc112bcd106bb1d82fef84fa632c6683401d13a6b66552

Observation 31f97f64-9b56-4d71-aa35-d99cc0f5a528 · outbound

This paper cites PointLLM: Empowering Large Language Models to Understand Point Clouds.

Generating 6DoF Object Manipulation Trajectories from Action Description in Egocentric Vision PointLLM: Empowering Large Language Models to Understand Point Clouds

Reference 97

Resolution
unresolved
no resolver link, observed 2026-08-07T11:03:01.866129Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:03:01.866129Z digest=sha256:503eee424c58b819717a6a1bf5cd4bbcb0f6b9d53f16e6d134d511706b6590a9

Observation 00399a2c-a072-43b7-a435-f32a7499ca06 · outbound

This paper cites Ulip- 2: Towards scalable multimodal pre-training for 3d under- standing.

Generating 6DoF Object Manipulation Trajectories from Action Description in Egocentric Vision Ulip- 2: Towards scalable multimodal pre-training for 3d under- standing

Reference 98

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:03:02.417177Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T11:03:01.868766Z digest=sha256:0fa797ed03189a262855cf718dfb6aeda49e67c66b776aab18801e0226509cf4

Observation d6fc5a06-7719-4580-ab73-e4b6970ab3ac · outbound

This paper cites Active object detection with knowledge aggregation and distillation from large models.

Generating 6DoF Object Manipulation Trajectories from Action Description in Egocentric Vision Active object detection with knowledge aggregation and distillation from large models

Reference 99

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:03:02.410018Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T11:03:01.871012Z digest=sha256:44165df374ee78d3e7bcd29ef40350c65323e35e0b4f596a24d77dabab1b47e6

Observation 6732de36-5603-4a24-8591-10bd181f752a · outbound

This paper cites Depth Anything: Unleashing the Power of Large-Scale Unlabeled Data.

Generating 6DoF Object Manipulation Trajectories from Action Description in Egocentric Vision Depth Anything: Unleashing the Power of Large-Scale Unlabeled Data

Reference 100

Resolution
unresolved
no resolver link, observed 2026-08-07T11:03:01.873084Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:03:01.873084Z digest=sha256:3249e60ba291614adb8fbe0e8c0d89f56f450979a43bfc7d4195d38ee992d2ee

Pith citing papers

Observation 8bda7ddf-e769-428a-b03a-0e6824903cce · inbound

MotionForesight: Re-purposing Video Models for Future 3D Scene-Flow Prediction cites this paper.

MotionForesight: Re-purposing Video Models for Future 3D Scene-Flow Prediction Generating 6DoF Object Manipulation Trajectories from Action Description in Egocentric Vision

Reference 6

Resolution
verified exact
local_arxiv, observed 2026-08-01T21:08:34.652209Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-01T21:08:03.385771Z digest=sha256:13d5e96bceec5c340236157c3efb825d007c1316ea5f0a24411d747cf2fe4d1e