Pith. sign in

Paper Citation Record · LEDGER

Generating 6DoF Object Manipulation Trajectories from Action Description in Egocentric Vision

As of 18 August 2026, this Paper Citation Record lists 100 of 111 outbound references and 1 inbound Pith citation observation for arXiv:2506.03605.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2506.03605 v1

Coverage vector

measured 100 of 111 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T11:03:01.873084Z

measured 101 of 101 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-18T06:34:40.430872+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-01T21:08:03.385771Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-08-05T02:28:24.338817Z

Reference resolution

100 of 111 outbound references displayed

  • verified exact0
  • verified fuzzy45
  • unresolved55
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

0
pith, observed 2026-08-05T02:28:24.338817Z

Outbound references

Observation 47410969-3493-4621-8539-833a4921c801 · outbound

This paper cites GPT-4 Technical Report.

Generating 6DoF Object Manipulation Trajectories from Action Description in Egocentric Vision GPT-4 Technical Report

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-07T11:03:01.639168Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:03:01.639168Z digest=sha256:b621d65db2896f6eb2679f7615da128b825c99d2ce62bf4a7cc2d8877205959b

Observation 0e0448d2-0fd0-46e8-b11f-0be3ee0c163c · outbound

This paper cites Affordances from human videos as a versatile representation for robotics.

Generating 6DoF Object Manipulation Trajectories from Action Description in Egocentric Vision Affordances from human videos as a versatile representation for robotics

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-07T11:03:01.643039Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:03:01.643039Z digest=sha256:ce91b77f5772e199d6f7d3cd5a8d2ca392e4dce7934d91344700732040b81cf0

Observation eacf354e-7547-4189-8c4f-a1aaa84966a9 · outbound

This paper cites Introducing HOT3D: An Egocentric Dataset for 3D Hand and Object Tracking.

Generating 6DoF Object Manipulation Trajectories from Action Description in Egocentric Vision Introducing HOT3D: An Egocentric Dataset for 3D Hand and Object Tracking

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-07T11:03:01.645783Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:03:01.645783Z digest=sha256:39082d3bafe32d0e790873910d8d0d358240f04e04e16a81c7d61a0869281cc1

Observation c5de6b31-54c3-43ed-8559-99d5c96e98dc · outbound

This paper cites METEOR: An auto- matic metric for MT evaluation with improved correlation with human judgments.

Generating 6DoF Object Manipulation Trajectories from Action Description in Egocentric Vision METEOR: An auto- matic metric for MT evaluation with improved correlation with human judgments

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-07T11:03:01.648655Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:03:01.648655Z digest=sha256:23d0857b608b7c93160d6710cd2712900fc995f92c752424c94d474845e3069a

Observation 11cb9d9e-ebb4-4334-ba79-b778064223d4 · outbound

This paper cites Uncertainty-aware state space transformer for egocentric 3d hand trajectory forecasting.

Generating 6DoF Object Manipulation Trajectories from Action Description in Egocentric Vision Uncertainty-aware state space transformer for egocentric 3d hand trajectory forecasting

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-07T11:03:01.651358Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:03:01.651358Z digest=sha256:e5b9dea5027cc50bed50320cfb2afc3f775554041277b4e13179194bda57c101

Observation de485560-fb69-47ec-9292-dd7e502dd5d2 · outbound

This paper cites ARKitscenes - a diverse real-world dataset for 3d indoor scene understanding using mobile RGB-d data.

Generating 6DoF Object Manipulation Trajectories from Action Description in Egocentric Vision ARKitscenes - a diverse real-world dataset for 3d indoor scene understanding using mobile RGB-d data

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-07T11:03:01.653992Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:03:01.653992Z digest=sha256:f40084f733ad3b054c3e60a1f8e19a272153483205be9a3bfcb714e8d98adf79

Observation b1cea614-4c66-407d-9791-5c25fa705b26 · outbound

This paper cites Yu, and Jianbo Shi.

Generating 6DoF Object Manipulation Trajectories from Action Description in Egocentric Vision Yu, and Jianbo Shi

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-07T11:03:01.656749Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:03:01.656749Z digest=sha256:4c8f753b57747888be58c99f2489c6c7c40af46fa6033b01125f005750cdb37f

Observation 92a0aeb6-a92e-4280-802d-537d3320f81e · outbound

This paper cites Kemp, and James Hays.

Generating 6DoF Object Manipulation Trajectories from Action Description in Egocentric Vision Kemp, and James Hays

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-07T11:03:01.659454Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:03:01.659454Z digest=sha256:fe9e0e0779d2e6b085e239edaec63b0fcfb20134ff314b10fbeceb9a5a147ef5

Observation b2979133-5d9f-4006-9856-5fe7c1fc25cc · outbound

This paper cites RT-1: Robotics Transformer for Real- World Control at Scale.

Generating 6DoF Object Manipulation Trajectories from Action Description in Egocentric Vision RT-1: Robotics Transformer for Real- World Control at Scale

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-07T11:03:01.662279Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:03:01.662279Z digest=sha256:eec7c27175901e76f79314b9e5c06c005dc15b6e23d0a7c239c7c736af3f1981

Observation 4c9c4bda-f8b5-4526-af9a-3b52d7b56fb1 · outbound

This paper cites Deep regression on manifolds: A 3d ro- tation case study.

Generating 6DoF Object Manipulation Trajectories from Action Description in Egocentric Vision Deep regression on manifolds: A 3d ro- tation case study

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-07T11:03:01.665148Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:03:01.665148Z digest=sha256:ea75e3dfc22003e8f0e7bf30d3e4839b8268f23f82507c24675453b2fd13c759

Observation 557b2e18-0eaf-43e4-ba88-bce57396ee90 · outbound

This paper cites Text2hoi: Text-guided 3d motion generation for hand-object interaction.

Generating 6DoF Object Manipulation Trajectories from Action Description in Egocentric Vision Text2hoi: Text-guided 3d motion generation for hand-object interaction

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-07T11:03:01.667377Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:03:01.667377Z digest=sha256:8c9b9204a33810f24d273a8a9775e33768a6cc860e0057897a3100badbc695e6

Observation 561ecdfc-d092-4cf9-9d8b-f12debce7731 · outbound

This paper cites Fleet, and Geoffrey Hinton.

Generating 6DoF Object Manipulation Trajectories from Action Description in Egocentric Vision Fleet, and Geoffrey Hinton

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-07T11:03:01.669846Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:03:01.669846Z digest=sha256:c7d14e22bea25928e27ad8584dbfb0871bc595259609ebee9748ee5e87eb1419

Observation ea0f3456-6039-4168-b018-0686df5c26ad · outbound

This paper cites Looking to relations for future trajectory forecast.

Generating 6DoF Object Manipulation Trajectories from Action Description in Egocentric Vision Looking to relations for future trajectory forecast

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-07T11:03:01.672243Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:03:01.672243Z digest=sha256:436340643d69e688616c7e5f97ccc53e24f75ad2f339e9f121a63b306e7738d7

Observation 7e55b631-8c15-421e-bff0-bcadd4487b6d · outbound

This paper cites Learning to act properly: Predicting and explain- ing affordances from images.

Generating 6DoF Object Manipulation Trajectories from Action Description in Egocentric Vision Learning to act properly: Predicting and explain- ing affordances from images

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-07T11:03:01.674396Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:03:01.674396Z digest=sha256:bcdc4fa2a0c0120ba27f02778208a359807826ccfbf3315c02b1f5eb65d629d1

Observation a68d4040-8d7d-4549-97c3-31865a1c710a · outbound

This paper cites Open x-embodiment: Robotic learning datasets and rt-x models: Open x- embodiment collaboration.

Generating 6DoF Object Manipulation Trajectories from Action Description in Egocentric Vision Open x-embodiment: Robotic learning datasets and rt-x models: Open x- embodiment collaboration

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-07T11:03:01.676590Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:03:01.676590Z digest=sha256:debbdec2d4ef39082efda97fd3a9496e1980e1e82179580720e49f24dc819959

Observation 27c3dced-da60-421b-bc08-b56860c1b583 · outbound

This paper cites Ganhand: Predicting human grasp affordances in multi-object scenes.

Generating 6DoF Object Manipulation Trajectories from Action Description in Egocentric Vision Ganhand: Predicting human grasp affordances in multi-object scenes

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-07T11:03:01.678794Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:03:01.678794Z digest=sha256:6b6f09f3b96efbc79d03b412db7d12e81c56d2776c4d68e48e26ecfa34302ad7

Observation 1fb4cb0e-aa1b-4380-a678-8667f639a872 · outbound

This paper cites Scaling egocentric vision: The epic- kitchens dataset.

Generating 6DoF Object Manipulation Trajectories from Action Description in Egocentric Vision Scaling egocentric vision: The epic- kitchens dataset

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-07T11:03:01.681123Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:03:01.681123Z digest=sha256:3ffb2a49b3e760980df34352ab6113183d75abfdb6506cdfb525651ecc59eb61

Observation 5b0e67e0-837d-4c2a-8850-75b5b89d204e · outbound

This paper cites Rescaling egocentric vision: Collection, pipeline and challenges for epic-kitchens-100.Interna- tional Journal of Computer Vision, 130(1):33–55, 2022.

Generating 6DoF Object Manipulation Trajectories from Action Description in Egocentric Vision Rescaling egocentric vision: Collection, pipeline and challenges for epic-kitchens-100.Interna- tional Journal of Computer Vision, 130(1):33–55, 2022

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-07T11:03:01.683615Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:03:01.683615Z digest=sha256:5997c0211cfb027788f191f45c102d61a845245cdbe1b544424029b266458a86

Observation de27e247-7e19-4ed7-b155-7edbf824ff29 · outbound

This paper cites Affor- dancenet: An end-to-end deep learning approach for object affordance detection.

Generating 6DoF Object Manipulation Trajectories from Action Description in Egocentric Vision Affor- dancenet: An end-to-end deep learning approach for object affordance detection

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-07T11:03:01.685866Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:03:01.685866Z digest=sha256:8e8f9f52372d3025dcf9daa509b4893796f3ea046bc6b1e19906cf84befcdc09

Observation cb20bff5-da7a-44e5-bcec-6300eccf0542 · outbound

This paper cites an unresolved cited work.

Generating 6DoF Object Manipulation Trajectories from Action Description in Egocentric Vision Unresolved cited work

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-07T11:03:01.688348Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:03:01.688348Z digest=sha256:f017e1be80ba9c520507a37a535099526aad7d97be78304ef9c22c3f3a9e834a

Observation f70df54e-48eb-4c63-b46c-d28fefc7c550 · outbound

This paper cites Project Aria: A New Tool for Egocentric Multi-Modal AI Research.

Generating 6DoF Object Manipulation Trajectories from Action Description in Egocentric Vision Project Aria: A New Tool for Egocentric Multi-Modal AI Research

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-07T11:03:01.690863Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:03:01.690863Z digest=sha256:1c4a7af081e852a04b314a6bf333de78b47dbdb454ba9fab4e78db60d2c9524c

Observation 4cafee9a-65f6-46df-be94-96f49ae2bf13 · outbound

This paper cites Egopat3dv2: Predicting 3d action target from 2d egocentric vision for human-robot interac- tion.

Generating 6DoF Object Manipulation Trajectories from Action Description in Egocentric Vision Egopat3dv2: Predicting 3d action target from 2d egocentric vision for human-robot interac- tion

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-07T11:03:01.693195Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:03:01.693195Z digest=sha256:f169721e00894ded6a4c0d0b3b28f140b384414ece02af170b580235b5d61241

Observation 4f40752f-fb0c-4820-8e37-cacb93de807b · outbound

This paper cites Fischler and Robert C.

Generating 6DoF Object Manipulation Trajectories from Action Description in Egocentric Vision Fischler and Robert C

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-07T11:03:01.695507Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:03:01.695507Z digest=sha256:706f11c11ecbfa8ec8fa57f4c603e49cbcb439f04ba573779365d02aef9853d0

Observation 822acb03-881c-4f2b-9c00-033718bdcf53 · outbound

This paper cites Zhao, and Chelsea Finn.

Generating 6DoF Object Manipulation Trajectories from Action Description in Egocentric Vision Zhao, and Chelsea Finn

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-07T11:03:01.697706Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:03:01.697706Z digest=sha256:67e5e7d5f36b637d783c0fc08da37d4c6c928299a99354e18a952e8605cad19d

Observation bf437845-6a3b-4c46-846b-5ddb1e278a48 · outbound

This paper cites What would you expect? anticipating egocentric actions with rolling-unrolling lstms and modality attention.

Generating 6DoF Object Manipulation Trajectories from Action Description in Egocentric Vision What would you expect? anticipating egocentric actions with rolling-unrolling lstms and modality attention

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-07T11:03:01.699904Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:03:01.699904Z digest=sha256:0d0633b84ea7a883e7707dd55d03348e445c17e70107c1be7adc8bdf50cbd007

Observation b43f8726-c20e-4eb5-b077-dd31dad3d841 · outbound

This paper cites Next-active-object predic- tion from egocentric videos.Journal of Visual Communi- cation and Image Representation, 49:401–411, 2017.

Generating 6DoF Object Manipulation Trajectories from Action Description in Egocentric Vision Next-active-object predic- tion from egocentric videos.Journal of Visual Communi- cation and Image Representation, 49:401–411, 2017

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-07T11:03:01.702296Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:03:01.702296Z digest=sha256:0af7937b31c057a588fd43766458132d4279fbb30fa900333cd9eb8e1b95ba86

Observation c23d645a-23b6-4211-ae47-869660f44c20 · outbound

This paper cites So predictable! continuous 3d hand trajectory prediction in virtual reality.

Generating 6DoF Object Manipulation Trajectories from Action Description in Egocentric Vision So predictable! continuous 3d hand trajectory prediction in virtual reality

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-07T11:03:01.705143Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:03:01.705143Z digest=sha256:f331d1b8e23a35b5834a5088e6cfa3155d7b49c83467c843752372a0c2757225

Observation df8fd2ba-d4cd-4342-a041-96d9b3477dc5 · outbound

This paper cites Transformer networks for trajectory forecasting.

Generating 6DoF Object Manipulation Trajectories from Action Description in Egocentric Vision Transformer networks for trajectory forecasting

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-07T11:03:01.707448Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:03:01.707448Z digest=sha256:bc997b460ab9c54dc5e3409d658ca10b71e837cded6884ab3b62d292c3a64742

Observation 38f75670-71f1-4db6-bf5a-88f0d6751f35 · outbound

This paper cites Ego4d: Around the world in 3,000 hours of egocentric video.

Generating 6DoF Object Manipulation Trajectories from Action Description in Egocentric Vision Ego4d: Around the world in 3,000 hours of egocentric video

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-07T11:03:01.709696Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:03:01.709696Z digest=sha256:b3afd9ae2b17ae80c1321d046bd8ef768c51333011d18b1f2af30b2ec8882947

Observation 8af6e787-429b-47bc-981c-a3128c1425d3 · outbound

This paper cites Ego-Exo4D: Understanding Skilled Human Activity from First- and Third-Person Perspectives.

Generating 6DoF Object Manipulation Trajectories from Action Description in Egocentric Vision Ego-Exo4D: Understanding Skilled Human Activity from First- and Third-Person Perspectives

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-07T11:03:01.711729Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:03:01.711729Z digest=sha256:0ac5bcf7b3235d53e7f6cc4852efbfc955099454ff1f1d5bb8a086b39600be64

Observation 21d1c34b-ed98-4cfd-8483-993549dd4e4a · outbound

This paper cites Handal: A dataset of real-world manipulable object cate- gories with pose annotations, affordances, and reconstruc- tions.

Generating 6DoF Object Manipulation Trajectories from Action Description in Egocentric Vision Handal: A dataset of real-world manipulable object cate- gories with pose annotations, affordances, and reconstruc- tions

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-07T11:03:01.714360Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:03:01.714360Z digest=sha256:7a93cfa948ff8f4f813c0feccc5f19cbe3c547f64ab2314edd07f441ce636482

Observation 931b5475-e2fb-436a-a2d1-d96678ec075c · outbound

This paper cites Honnotate: A method for 3d annotation of hand and object poses.

Generating 6DoF Object Manipulation Trajectories from Action Description in Egocentric Vision Honnotate: A method for 3d annotation of hand and object poses

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-07T11:03:01.716689Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:03:01.716689Z digest=sha256:1033816d3892048def2e773c384bf2b5375cec2bf4333451730bbfefbe95720a

Observation cb6f7cc7-ddf8-4477-837d-b16ebfd6c1fe · outbound

This paper cites Ego3dt: Tracking every 3d object in ego-centric videos.

Generating 6DoF Object Manipulation Trajectories from Action Description in Egocentric Vision Ego3dt: Tracking every 3d object in ego-centric videos

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-07T11:03:01.718900Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:03:01.718900Z digest=sha256:781694c94c211463ebda4223f6c7ed6c973d5431feff003e798493a7a0f9d77d

Observation adc330e8-6101-4ca2-9e30-b0f7e09ed2a3 · outbound

This paper cites Onepose++: Keypoint-free one- shot object pose estimation without cad models.

Generating 6DoF Object Manipulation Trajectories from Action Description in Egocentric Vision Onepose++: Keypoint-free one- shot object pose estimation without cad models

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-07T11:03:01.721146Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:03:01.721146Z digest=sha256:e08a79d24116959edc4c7d9971f5eaf7149c55dbf67ce4d4eb5603aa95f797f2

Observation b4bc772e-e112-43ec-8d45-b39e3fa6dde2 · outbound

This paper cites Pvn3d: A deep point-wise 3d keypoints voting network for 6dof pose estimation.

Generating 6DoF Object Manipulation Trajectories from Action Description in Egocentric Vision Pvn3d: A deep point-wise 3d keypoints voting network for 6dof pose estimation

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-07T11:03:01.723553Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:03:01.723553Z digest=sha256:92e12bf71469877cf464036d6df0da46ab5f13aea2c441c711894e7b113deebc

Observation d14f3dfe-27d0-4714-aa2c-d098d497726d · outbound

This paper cites Fs6d: Few-shot 6d pose estimation of novel objects.

Generating 6DoF Object Manipulation Trajectories from Action Description in Egocentric Vision Fs6d: Few-shot 6d pose estimation of novel objects

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:03:02.834333Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T11:03:01.725649Z digest=sha256:7d3883f266d31d0119420b455d78c623c3c976107ff8fe5fd68b283af4258f74

Observation fe9080e1-5070-4185-856b-56a5358a222d · outbound

This paper cites The curious case of neural text degeneration.

Generating 6DoF Object Manipulation Trajectories from Action Description in Egocentric Vision The curious case of neural text degeneration

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:03:02.827197Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T11:03:01.727761Z digest=sha256:6a407148f01336ce942a1e7ed2f3a834dc680dae0592f55a353ce1c05f3d0fed

Observation 17033dfd-1bed-4186-9d77-dc94c135f42a · outbound

This paper cites 3d-llm: Inject- ing the 3d world into large language models.

Generating 6DoF Object Manipulation Trajectories from Action Description in Egocentric Vision 3d-llm: Inject- ing the 3d world into large language models

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:03:02.820454Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T11:03:01.729959Z digest=sha256:c49dc51f993ed58e95ccd7f0d84d8cae6020cadc005cacc2e6d7500a4e4c579b

Observation 04037fe9-cbc4-4c92-917d-e2c1e3160f1b · outbound

This paper cites LITA: Language Instructed Temporal-Localization Assistant.

Generating 6DoF Object Manipulation Trajectories from Action Description in Egocentric Vision LITA: Language Instructed Temporal-Localization Assistant

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-07T11:03:01.732235Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:03:01.732235Z digest=sha256:1810289516206825f31d31552c8a7fdfc4de06863d5b30f6ffada1f0e696c13c

Observation 764af670-8aab-4d8c-bf47-021987343d08 · outbound

This paper cites V oxposer: Composable 3d value maps for robotic manipulation with language models.

Generating 6DoF Object Manipulation Trajectories from Action Description in Egocentric Vision V oxposer: Composable 3d value maps for robotic manipulation with language models

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:03:02.813619Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T11:03:01.734633Z digest=sha256:567d2c5ddea68f6ca1ca831967d03e113f51f1a8d1c72af836981996d093a259

Observation 2ef80ba7-fecf-4573-96ec-35d4f8b57f4e · outbound

This paper cites Technical Report for Ego4D Long Term Action Anticipation Challenge 2023.

Generating 6DoF Object Manipulation Trajectories from Action Description in Egocentric Vision Technical Report for Ego4D Long Term Action Anticipation Challenge 2023

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-07T11:03:01.737020Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:03:01.737020Z digest=sha256:05afc5abd6d31067078ed8c716cb0a0f82926b5b3ddfcbf01da8b73aa8c9307e

Observation f30ae56d-c238-4ccf-bb09-e808ef756a06 · outbound

This paper cites Jacobs, Michael I.

Generating 6DoF Object Manipulation Trajectories from Action Description in Egocentric Vision Jacobs, Michael I

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:03:02.806848Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T11:03:01.739593Z digest=sha256:395a838c2eb981e414165e63ce8f3478c42a2470e84896ab09c7cfa6a4f82ca8

Observation edbbc3bc-09a9-45cd-8a51-796d0fa0dfab · outbound

This paper cites Jordan and R.A.

Generating 6DoF Object Manipulation Trajectories from Action Description in Egocentric Vision Jordan and R.A

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:03:02.800121Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T11:03:01.741741Z digest=sha256:e25a0388a74cca88df112db9a0620b6d270575733b2b80cd8d72e80336da4539

Observation 6771cede-c617-48fe-9099-84ac18712688 · outbound

This paper cites OpenVLA: An Open-Source Vision-Language-Action Model.

Generating 6DoF Object Manipulation Trajectories from Action Description in Egocentric Vision OpenVLA: An Open-Source Vision-Language-Action Model

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-07T11:03:01.743937Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:03:01.743937Z digest=sha256:2c53789be0d6bb598ec3761a01ecaed00094ca3bfafec0aefa0cac24976559ee

Observation c03a6023-fbd9-4c1a-b77a-4f4f58c3a29a · outbound

This paper cites Koppula and Ashutosh Saxena.

Generating 6DoF Object Manipulation Trajectories from Action Description in Egocentric Vision Koppula and Ashutosh Saxena

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:03:02.792891Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T11:03:01.746343Z digest=sha256:5784c63d9903cdfb839ed7de67a407d40ed7b912563eefc04f4c1b5610bd3da4

Observation 7746be9a-727f-4292-9c42-34e4e355dd73 · outbound

This paper cites H2o: Two hands manipulating objects for first person interaction recognition.

Generating 6DoF Object Manipulation Trajectories from Action Description in Egocentric Vision H2o: Two hands manipulating objects for first person interaction recognition

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:03:02.785551Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T11:03:01.748950Z digest=sha256:1c4187a2a1bb1f07670939d3409c5715c4ecc7ba70857c9a4ce122df39333cee

Observation fd9ce56b-5902-45a9-9686-86d7318c9ae2 · outbound

This paper cites Choy, Philip H.

Generating 6DoF Object Manipulation Trajectories from Action Description in Egocentric Vision Choy, Philip H

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:03:02.778345Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T11:03:01.751582Z digest=sha256:3947be7cf55102b153c94674cd7553020ad524cc5f573061a7b3dbbf44fc6bc4

Observation 5d5f9080-3912-49cf-9ac3-2e07b2aebf5d · outbound

This paper cites Locate: Localize and transfer object parts for weakly supervised affordance grounding.

Generating 6DoF Object Manipulation Trajectories from Action Description in Egocentric Vision Locate: Localize and transfer object parts for weakly supervised affordance grounding

Reference 48

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:03:02.771210Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T11:03:01.753872Z digest=sha256:5f64f0c3b24a490e1a5671fe5a1f5a75aec8bc302f584d978411455bb48ddf6e

Observation 6891d889-8271-47ca-9ab6-a7ea2e73b333 · outbound

This paper cites Learning precise affordances from ego- centric videos for robotic manipulation.arxiv preprint arXiv:2408.10123, 2024.

Generating 6DoF Object Manipulation Trajectories from Action Description in Egocentric Vision Learning precise affordances from ego- centric videos for robotic manipulation.arxiv preprint arXiv:2408.10123, 2024

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-07T11:03:01.755952Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:03:01.755952Z digest=sha256:0884a702ffa534735dee93b79a19a6f9b5111ff3c86a3ef7cacf54baf8094135

Observation 2a635c79-6523-4826-af66-c6660f6a560c · outbound

This paper cites BLIP-2: Bootstrapping language-image pre-training with frozen image encoders and large language models.

Generating 6DoF Object Manipulation Trajectories from Action Description in Egocentric Vision BLIP-2: Bootstrapping language-image pre-training with frozen image encoders and large language models

Reference 50

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:03:02.763851Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T11:03:01.758183Z digest=sha256:8553c12dd5da398ac51ec4cd609569954a5a74f1f4dfdb3e8e714ea7c4461fa9

Observation ccb09b42-4bb9-4ce6-abcb-4a4662bc6d85 · outbound

This paper cites an unresolved cited work.

Generating 6DoF Object Manipulation Trajectories from Action Description in Egocentric Vision Unresolved cited work

Reference 51

Resolution
unresolved
raw_fallback, observed 2026-08-07T11:03:02.756194Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T11:03:01.760603Z digest=sha256:8e0cc61cfa551b82afc47e7f58acb6cb9fc188d43ed4a6c5cec11180550e2fbe

Observation 3837d5c2-7a9c-4e69-941e-4ace089e592c · outbound

This paper cites Deepim: Deep iterative matching for 6d pose estimation.

Generating 6DoF Object Manipulation Trajectories from Action Description in Egocentric Vision Deepim: Deep iterative matching for 6d pose estimation

Reference 52

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:03:02.749185Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T11:03:01.762807Z digest=sha256:c63c3363703a86125d60118c5936d3406f7c9b8b956e19ac5248cca4fb57e8f0

Observation e991d2d3-bc65-4b85-998d-397973cb58d8 · outbound

This paper cites Egocentric predic- tion of action target in 3d.

Generating 6DoF Object Manipulation Trajectories from Action Description in Egocentric Vision Egocentric predic- tion of action target in 3d

Reference 53

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:03:02.741266Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T11:03:01.765023Z digest=sha256:f06e6f4712f49a96629aae5fbbe9672a4a2ddee2ddba0470a39bbf8f591eabc2

Observation 4b58a739-6ba0-40d7-b793-ad66de36a381 · outbound

This paper cites Vila: On pre-training for vi- sual language models.

Generating 6DoF Object Manipulation Trajectories from Action Description in Egocentric Vision Vila: On pre-training for vi- sual language models

Reference 54

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:03:02.733369Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T11:03:01.767024Z digest=sha256:84a467c49f8fcba9504b6443d2466da147ccddbd8eb67391d7802fb3862aef6d

Observation 1aaac6ff-ea91-4329-a438-0c2fedd866c8 · outbound

This paper cites Visual instruction tuning.

Generating 6DoF Object Manipulation Trajectories from Action Description in Egocentric Vision Visual instruction tuning

Reference 55

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:03:02.725673Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T11:03:01.769093Z digest=sha256:b6c805609b6782bd4b93be471a15b4ced4d078e2db08f814bd1356a7db69e1dd

Observation c9bba91b-93f4-4abe-8075-a3ad36481ef3 · outbound

This paper cites Improved baselines with visual instruction tuning.

Generating 6DoF Object Manipulation Trajectories from Action Description in Egocentric Vision Improved baselines with visual instruction tuning

Reference 56

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:03:02.718489Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T11:03:01.771408Z digest=sha256:d66e00455a7798f701e6ccc5e1d88daa73c0617d15c2e0bdf5b7f74a4dded120

Observation 95b62291-4839-402c-95b4-e81f6f958610 · outbound

This paper cites an unresolved cited work.

Generating 6DoF Object Manipulation Trajectories from Action Description in Egocentric Vision Unresolved cited work

Reference 57

Resolution
unresolved
raw_fallback, observed 2026-08-07T11:03:02.711058Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T11:03:01.773689Z digest=sha256:85f1040f8ab1548f26277803491f24b7778f1c33a0c9297c3361db0581479b23

Observation 89577c0a-d18c-4960-8da4-d952bdd5d7ef · outbound

This paper cites Joint hand motion and interaction hotspots prediction from egocentric videos.

Generating 6DoF Object Manipulation Trajectories from Action Description in Egocentric Vision Joint hand motion and interaction hotspots prediction from egocentric videos

Reference 58

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:03:02.703822Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T11:03:01.776268Z digest=sha256:6ad789758400fd4a75e49635ae9a4d9a86655877e1100f91c20c7b975875d526

Observation 447f20c7-76fc-47ae-9800-9c979c79da83 · outbound

This paper cites Grounding DINO: Marrying DINO with Grounded Pre-Training for Open-Set Object Detection.

Generating 6DoF Object Manipulation Trajectories from Action Description in Egocentric Vision Grounding DINO: Marrying DINO with Grounded Pre-Training for Open-Set Object Detection

Reference 59

Resolution
unresolved
no resolver link, observed 2026-08-07T11:03:01.778564Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:03:01.778564Z digest=sha256:0bf369ab09ccc5c974c6d3b9afead4b7320a73752a8b14c30f6aa54321914002

Observation 023e3407-0794-4240-b2a2-df951c391638 · outbound

This paper cites Hoi4d: A 4d egocentric dataset for category-level human-object interaction.

Generating 6DoF Object Manipulation Trajectories from Action Description in Egocentric Vision Hoi4d: A 4d egocentric dataset for category-level human-object interaction

Reference 60

Resolution
unresolved
no resolver link, observed 2026-08-07T11:03:01.781185Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:03:01.781185Z digest=sha256:6cf51a77ada0aa645383ebce61e827fc1f4f9d4a7a83824e090fcdb434d6b0f5

Observation e08fbb80-4ef6-4756-93f0-7476aee864c0 · outbound

This paper cites Decoupled weight decay regularization.

Generating 6DoF Object Manipulation Trajectories from Action Description in Egocentric Vision Decoupled weight decay regularization

Reference 61

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:03:02.691148Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T11:03:01.783372Z digest=sha256:0e8193ff241f62008b738cd9fdc860dbaa9bb75c97bfe7e6baace36b6afef2da

Observation 930debae-048d-41d0-a0eb-29c7346224a7 · outbound

This paper cites Phrase-based affordance detection via cyclic bi- lateral interaction.IEEE Transactions on Artificial Intelli- gence, 4(5):1186–1198, 2023.

Generating 6DoF Object Manipulation Trajectories from Action Description in Egocentric Vision Phrase-based affordance detection via cyclic bi- lateral interaction.IEEE Transactions on Artificial Intelli- gence, 4(5):1186–1198, 2023

Reference 62

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:03:02.684074Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T11:03:01.785682Z digest=sha256:c20a6784c833b3bc90655006ed1f40183c39adcc2d391414205ef6dac7363fb8

Observation f15860b1-d882-4991-b12c-21228494844b · outbound

This paper cites Madiff: Motion-aware mamba diffusion models for hand trajectory prediction on egocentric videos.

Generating 6DoF Object Manipulation Trajectories from Action Description in Egocentric Vision Madiff: Motion-aware mamba diffusion models for hand trajectory prediction on egocentric videos

Reference 63

Resolution
unresolved
no resolver link, observed 2026-08-07T11:03:01.787848Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:03:01.787848Z digest=sha256:c5e6138e22706f9d0301c062583cb3257c59a6d4ff1bf7fdc9db4b425a6129e2

Observation 2026da06-56bc-47f1-846b-17d4caab190a · outbound

This paper cites Diff-ip2d: Diffusion-based hand-object interac- tion prediction on egocentric videos.arXiv preprint arXiv:2405.04370, 2024.

Generating 6DoF Object Manipulation Trajectories from Action Description in Egocentric Vision Diff-ip2d: Diffusion-based hand-object interac- tion prediction on egocentric videos.arXiv preprint arXiv:2405.04370, 2024

Reference 64

Resolution
unresolved
no resolver link, observed 2026-08-07T11:03:01.789958Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:03:01.789958Z digest=sha256:a25215a1539a2b068e20251d02897de2e8adc4beb31c5558b62470d27f1631c8

Observation 5153fdae-f7e7-4702-99b5-59a4e0fca411 · outbound

This paper cites Quest 3, 2023.

Generating 6DoF Object Manipulation Trajectories from Action Description in Egocentric Vision Quest 3, 2023

Reference 65

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:03:02.676435Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T11:03:01.792768Z digest=sha256:e91955eea763b872c81fc8bddbd04e9572495d2422a53c3daa2ad17ecb1d472d

Observation 819c8625-4a37-4003-b1c7-e15ea6411a84 · outbound

This paper cites Leveraging the present to anticipate the future in videos.

Generating 6DoF Object Manipulation Trajectories from Action Description in Egocentric Vision Leveraging the present to anticipate the future in videos

Reference 66

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:03:02.668973Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T11:03:01.795177Z digest=sha256:421c5b450101044753c66eba82ea10863d01f64f140fbeb766614860d48b0bf3

Observation 1102a049-928a-4b00-8f73-b05afd3ba384 · outbound

This paper cites RT-Affordance: Affordances are Versatile Intermediate Representations for Robot Manipulation.

Generating 6DoF Object Manipulation Trajectories from Action Description in Egocentric Vision RT-Affordance: Affordances are Versatile Intermediate Representations for Robot Manipulation

Reference 67

Resolution
unresolved
no resolver link, observed 2026-08-07T11:03:01.797316Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:03:01.797316Z digest=sha256:17614bcd7246c980ffb7fe4f4dc604be2d34bafbc1dd9ec7db6c31c5dac354f2

Observation d6df2c9a-dc7a-42bc-8c84-71a04c76e048 · outbound

This paper cites Open-vocabulary affordance detection in 3d point clouds.

Generating 6DoF Object Manipulation Trajectories from Action Description in Egocentric Vision Open-vocabulary affordance detection in 3d point clouds

Reference 68

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:03:02.593310Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T11:03:01.799828Z digest=sha256:cedb0f972a4a1feab356cf94ec7608f1a93ad57d4890f98227b970d6f09514b1

Observation 8c71b377-7abe-4271-8ff7-7e4f1e62d16e · outbound

This paper cites Bleu: a method for automatic evaluation of machine translation.

Generating 6DoF Object Manipulation Trajectories from Action Description in Egocentric Vision Bleu: a method for automatic evaluation of machine translation

Reference 69

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:03:02.585613Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T11:03:01.802108Z digest=sha256:12e90de7738a42d280400999f6962fb51f3fe7c338a5fa9d21a153aff8de4753

Observation 42a10d2e-0ed0-4b9e-a118-10202a618bf0 · outbound

This paper cites Colored point cloud registration revisited.

Generating 6DoF Object Manipulation Trajectories from Action Description in Egocentric Vision Colored point cloud registration revisited

Reference 70

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:03:02.578279Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T11:03:01.804235Z digest=sha256:ad156b101163c3a4436b3737a96bde2be566b798811abca85543ec9ca0cca0c7

Observation 7b1dbff0-fffd-40ef-92ba-84967b0d8a59 · outbound

This paper cites Pix2pose: Pixel-wise coordinate regression of objects for 6d pose es- timation.

Generating 6DoF Object Manipulation Trajectories from Action Description in Egocentric Vision Pix2pose: Pixel-wise coordinate regression of objects for 6d pose es- timation

Reference 71

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:03:02.570682Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T11:03:01.806379Z digest=sha256:aef395a19899bc4d3b5a41a422f692184fdfa981d9b02f8d4375c82f13dccf59

Observation 8aab98a6-a389-4042-a712-40a55c6136f9 · outbound

This paper cites GloVe: Global vectors for word representation.

Generating 6DoF Object Manipulation Trajectories from Action Description in Egocentric Vision GloVe: Global vectors for word representation

Reference 72

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:03:02.561159Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T11:03:01.808650Z digest=sha256:805d8210e431e527dc525fafc80cd4fe35c15ef35f5013c7521e88ca5078292c

Observation 27344cf1-4e69-451c-be86-c28f395af6e0 · outbound

This paper cites Detecting activities of daily living in first-person camera views.

Generating 6DoF Object Manipulation Trajectories from Action Description in Egocentric Vision Detecting activities of daily living in first-person camera views

Reference 73

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:03:02.553759Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T11:03:01.810884Z digest=sha256:27429df2553729c286722e0b9e49eb2197dedf1cb53db6294cb752e8d743e313

Observation 5ea1d6b9-698f-42d3-8f70-1c24eae0f2bc · outbound

This paper cites Spatial Cognition from Egocentric Video: Out of Sight, Not Out of Mind.

Generating 6DoF Object Manipulation Trajectories from Action Description in Egocentric Vision Spatial Cognition from Egocentric Video: Out of Sight, Not Out of Mind

Reference 74

Resolution
unresolved
no resolver link, observed 2026-08-07T11:03:01.813217Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:03:01.813217Z digest=sha256:d6e362cb26ab29b5a8fb0ac5b56ed9e417e6ddb84884098f08a7c8ae9c1e806b

Observation fa4af39c-dada-4701-bfa9-78b139a080aa · outbound

This paper cites Learning transferable visual models from natural language supervision.

Generating 6DoF Object Manipulation Trajectories from Action Description in Egocentric Vision Learning transferable visual models from natural language supervision

Reference 75

Resolution
unresolved
no resolver link, observed 2026-08-07T11:03:01.815569Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:03:01.815569Z digest=sha256:34fd97ae51670cea1d83db77ae600dae5172969176bc10ec93a7f2ad50c39caf

Observation ed255f1f-fd84-435a-8580-9eb8727f6062 · outbound

This paper cites Grounded SAM: Assembling Open-World Models for Diverse Visual Tasks.

Generating 6DoF Object Manipulation Trajectories from Action Description in Egocentric Vision Grounded SAM: Assembling Open-World Models for Diverse Visual Tasks

Reference 76

Resolution
unresolved
no resolver link, observed 2026-08-07T11:03:01.817785Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:03:01.817785Z digest=sha256:fff4bc8a16e9591b7cd9c17251c00ff9093cbf4bef15e3bea0c61983e2f08407

Observation 81922796-4871-435b-9a34-03e8d7c49c91 · outbound

This paper cites Action scene graphs for long-form understanding of egocentric videos.

Generating 6DoF Object Manipulation Trajectories from Action Description in Egocentric Vision Action scene graphs for long-form understanding of egocentric videos

Reference 77

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:03:02.541814Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T11:03:01.820214Z digest=sha256:8a33c35f9413f510b611ed4f7ab6e91a713f7f344ff7336f0787a26e16a85032

Observation 14504436-79ae-42d3-9260-94546030c3fc · outbound

This paper cites Fast point feature histograms (fpfh) for 3d registration.

Generating 6DoF Object Manipulation Trajectories from Action Description in Egocentric Vision Fast point feature histograms (fpfh) for 3d registration

Reference 78

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:03:02.534947Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T11:03:01.823089Z digest=sha256:dfa0692e4a5c256f93a2a49c60add40d36a3522736087e7fc95f3fa1a2428632

Observation ade2407b-596b-4b88-8b2d-c95a416708f2 · outbound

This paper cites What object should i use? - task driven object detection.

Generating 6DoF Object Manipulation Trajectories from Action Description in Egocentric Vision What object should i use? - task driven object detection

Reference 79

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:03:02.527767Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T11:03:01.825425Z digest=sha256:a8681c2b7c7ebdf111effc8f7c14f896bf6caf2e44f4945cf6eb97630b34a267

Observation 1ea84205-78c7-42aa-a7b9-ebb01551aaca · outbound

This paper cites As- sembly101: A large-scale multi-view video dataset for un- derstanding procedural activities.

Generating 6DoF Object Manipulation Trajectories from Action Description in Egocentric Vision As- sembly101: A large-scale multi-view video dataset for un- derstanding procedural activities

Reference 80

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:03:02.520789Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T11:03:01.827561Z digest=sha256:da2a4aefb39be4f856104ba043fbfea243f1e992e87859c64668c1e655c83f09

Observation 02e18a5e-1626-432f-9be8-acb4e358afcc · outbound

This paper cites an unresolved cited work.

Generating 6DoF Object Manipulation Trajectories from Action Description in Egocentric Vision Unresolved cited work

Reference 81

Resolution
unresolved
raw_fallback, observed 2026-08-07T11:03:02.513616Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T11:03:01.829880Z digest=sha256:79d1449a38f7a3414394aaef5ccecad413e12b1b930ea0492b451d3cec22da15

Observation 03be063f-c5ca-4042-9db8-e65b21cf0d39 · outbound

This paper cites Ego4d goal-step: Toward hierarchical understanding of procedural activities.

Generating 6DoF Object Manipulation Trajectories from Action Description in Egocentric Vision Ego4d goal-step: Toward hierarchical understanding of procedural activities

Reference 82

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:03:02.506891Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T11:03:01.832360Z digest=sha256:09b340dec195f57f18377695aba3c1842c260f9e0920c19fdd43131589586011

Observation d5f21668-f04e-48dd-a9d6-ae5ca92bef9b · outbound

This paper cites Onepose: One-shot object pose estimation without cad models.

Generating 6DoF Object Manipulation Trajectories from Action Description in Egocentric Vision Onepose: One-shot object pose estimation without cad models

Reference 83

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:03:02.499917Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T11:03:01.834662Z digest=sha256:ada2f176e6764878d3d1e0886933bdc30072404143a746f10427c5c8bc413f2a

Observation fba8b8e1-790e-440e-895c-388bb1383650 · outbound

This paper cites an unresolved cited work.

Generating 6DoF Object Manipulation Trajectories from Action Description in Egocentric Vision Unresolved cited work

Reference 84

Resolution
unresolved
raw_fallback, observed 2026-08-07T11:03:02.493194Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T11:03:01.836944Z digest=sha256:c80b1613403eef9511e6b1c3ca08606eef656583a45107a7c91581a7203a5aad

Observation 57486b92-00fb-424a-868f-721d11782568 · outbound

This paper cites MiniGPT-3D: Efficiently Aligning 3D Point Clouds with Large Language Models using 2D Priors.

Generating 6DoF Object Manipulation Trajectories from Action Description in Egocentric Vision MiniGPT-3D: Efficiently Aligning 3D Point Clouds with Large Language Models using 2D Priors

Reference 85

Resolution
unresolved
no resolver link, observed 2026-08-07T11:03:01.839325Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:03:01.839325Z digest=sha256:4983fdcf266430b38e934bb98a28d945477a4cd64442ffe188ac4d883fa4fd5a

Observation d8c0f507-059b-4f64-a679-6558183364cf · outbound

This paper cites Leveraging next-active ob- jects for context-aware anticipation in egocentric videos.

Generating 6DoF Object Manipulation Trajectories from Action Description in Egocentric Vision Leveraging next-active ob- jects for context-aware anticipation in egocentric videos

Reference 86

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:03:02.486631Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T11:03:01.841611Z digest=sha256:80798281a47808cbc3d7a97efa844950d25eb721854d57b35b56719db004087d

Observation ad808590-b6e7-45c3-bb81-048eb365dfbf · outbound

This paper cites LLaMA: Open and Efficient Foundation Language Models.

Generating 6DoF Object Manipulation Trajectories from Action Description in Egocentric Vision LLaMA: Open and Efficient Foundation Language Models

Reference 87

Resolution
unresolved
no resolver link, observed 2026-08-07T11:03:01.843780Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:03:01.843780Z digest=sha256:f958ab69e24cec254d3c890dc48f101e730efc06aeb8e11d6730a1777996811d

Observation 4e70f3ce-e83f-4137-9b97-2d60b237b59d · outbound

This paper cites Epic fields: Marrying 3d geometry and video un- derstanding.

Generating 6DoF Object Manipulation Trajectories from Action Description in Egocentric Vision Epic fields: Marrying 3d geometry and video un- derstanding

Reference 88

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:03:02.479483Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T11:03:01.845967Z digest=sha256:d62b17f3bdb0c9e48a56cba35dc040bd4192bdd040777fd9d6a23bd69f37f665

Observation dfc00f89-07c2-4bea-8149-091d1f37e093 · outbound

This paper cites Attention is all you need.

Generating 6DoF Object Manipulation Trajectories from Action Description in Egocentric Vision Attention is all you need

Reference 89

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:03:02.472614Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T11:03:01.848178Z digest=sha256:574c2ca6840cbb1c5fbd4ff3a4de2e62b18d1c544f95225b285005407b253ac1

Observation 765a3edd-8ca3-45de-94f2-35cdd54d589b · outbound

This paper cites Omnivid: A gener- ative framework for universal video understanding.

Generating 6DoF Object Manipulation Trajectories from Action Description in Egocentric Vision Omnivid: A gener- ative framework for universal video understanding

Reference 90

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:03:02.465495Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T11:03:01.850456Z digest=sha256:26bf526a419d4a74a5f6071f859bf7c29f1be3ed52ed6f63069d1b78647bca20

Observation e873cb05-ef21-4921-9033-1513b53c6e49 · outbound

This paper cites OFA: Unifying architectures, tasks, and modalities through a simple sequence-to-sequence learning framework.

Generating 6DoF Object Manipulation Trajectories from Action Description in Egocentric Vision OFA: Unifying architectures, tasks, and modalities through a simple sequence-to-sequence learning framework

Reference 91

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:03:02.458516Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T11:03:01.852711Z digest=sha256:496823cf336823830082cce7822e80f9dbf6b2c859a00ab58fd498d98312cae7

Observation ed3fb1cf-8abe-4e6b-951d-0cf6523a988b · outbound

This paper cites Dust3r: Geometric 3d vision made easy.

Generating 6DoF Object Manipulation Trajectories from Action Description in Egocentric Vision Dust3r: Geometric 3d vision made easy

Reference 92

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:03:02.451591Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T11:03:01.854962Z digest=sha256:90d6a55c689165d9bda1b15b5ba6f3831054f7347eb899f0344091b2e1653944

Observation ad6ba276-8925-4a1c-820e-bd952f9dd376 · outbound

This paper cites Holoassist: an egocentric human interaction dataset for interactive ai assistants in the real world.

Generating 6DoF Object Manipulation Trajectories from Action Description in Egocentric Vision Holoassist: an egocentric human interaction dataset for interactive ai assistants in the real world

Reference 93

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:03:02.444616Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T11:03:01.857232Z digest=sha256:ff40810aecb1c08fc6b06320a681e3c90f90555adb06fb5f867c88ec6ad438b4

Observation f21af856-db0f-4b4f-8b68-447966d98669 · outbound

This paper cites Foundationpose: Unified 6d pose estimation and tracking of novel objects.

Generating 6DoF Object Manipulation Trajectories from Action Description in Egocentric Vision Foundationpose: Unified 6d pose estimation and tracking of novel objects

Reference 94

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:03:02.437955Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T11:03:01.859550Z digest=sha256:445fb293ce6f10422f8c53c606778c69d2ef7813c5e1cc27ed67cd7b6abb89a2

Observation d361858b-5ecf-4bad-b00d-82f773432897 · outbound

This paper cites Learning descriptors for object recognition and 3d pose estimation.

Generating 6DoF Object Manipulation Trajectories from Action Description in Egocentric Vision Learning descriptors for object recognition and 3d pose estimation

Reference 95

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:03:02.430859Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T11:03:01.861969Z digest=sha256:79bb755bc1131c2cf499c9ab0ee440718c9e7b590e5bd59a72e5574b14887ce7

Observation 9ec32dc1-f846-42df-8dda-4e1cd383b105 · outbound

This paper cites Spatialtracker: Tracking any 2d pixels in 3d space.

Generating 6DoF Object Manipulation Trajectories from Action Description in Egocentric Vision Spatialtracker: Tracking any 2d pixels in 3d space

Reference 96

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:03:02.423930Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T11:03:01.864022Z digest=sha256:f4da3f6bf8f09e653a8c9153267dd09462e63f39b7c0a9ee125935ae120e445a

Observation 31f97f64-9b56-4d71-aa35-d99cc0f5a528 · outbound

This paper cites PointLLM: Empowering Large Language Models to Understand Point Clouds.

Generating 6DoF Object Manipulation Trajectories from Action Description in Egocentric Vision PointLLM: Empowering Large Language Models to Understand Point Clouds

Reference 97

Resolution
unresolved
no resolver link, observed 2026-08-07T11:03:01.866129Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:03:01.866129Z digest=sha256:651c7f45a8ef91854c8cc8f0311b0bd57287d31bf72af5f9703b248ef065af97

Observation 00399a2c-a072-43b7-a435-f32a7499ca06 · outbound

This paper cites Ulip- 2: Towards scalable multimodal pre-training for 3d under- standing.

Generating 6DoF Object Manipulation Trajectories from Action Description in Egocentric Vision Ulip- 2: Towards scalable multimodal pre-training for 3d under- standing

Reference 98

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:03:02.417177Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T11:03:01.868766Z digest=sha256:140fa3818f829c491439a9215ee6948763e9f3fd2ba4f77a0eefba1402e1d7b1

Observation d6fc5a06-7719-4580-ab73-e4b6970ab3ac · outbound

This paper cites Active object detection with knowledge aggregation and distillation from large models.

Generating 6DoF Object Manipulation Trajectories from Action Description in Egocentric Vision Active object detection with knowledge aggregation and distillation from large models

Reference 99

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:03:02.410018Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T11:03:01.871012Z digest=sha256:a25e21961e46f335c5138205bf9745d7da4c63ea37a0d09b01809176301d5c99

Observation 6732de36-5603-4a24-8591-10bd181f752a · outbound

This paper cites Depth Anything: Unleashing the Power of Large-Scale Unlabeled Data.

Generating 6DoF Object Manipulation Trajectories from Action Description in Egocentric Vision Depth Anything: Unleashing the Power of Large-Scale Unlabeled Data

Reference 100

Resolution
unresolved
no resolver link, observed 2026-08-07T11:03:01.873084Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:03:01.873084Z digest=sha256:f2177ca9c39580e9a18cf7f384a9d2e0922879b55f52817260fded5d61070932

Pith citing papers

Observation 8bda7ddf-e769-428a-b03a-0e6824903cce · inbound

MotionForesight: Re-purposing Video Models for Future 3D Scene-Flow Prediction cites this paper.

MotionForesight: Re-purposing Video Models for Future 3D Scene-Flow Prediction Generating 6DoF Object Manipulation Trajectories from Action Description in Egocentric Vision

Reference 6

Resolution
verified exact
local_arxiv, observed 2026-08-01T21:08:34.652209Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-01T21:08:03.385771Z digest=sha256:ebdb81edec201c85a9ab89f6fa10b741e150c8be8cfb14aa09506641f6447141