Pith. sign in

Paper Citation Record · LEDGER

Conditional Object-Centric Learning from Video

As of 14 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 19 inbound Pith citation observations for arXiv:2111.12594.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2111.12594 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 19 of 19 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-14T06:32:32.682623+00:00

measured 19 of 19 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-11T18:33:51.430960Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-03T10:48:02.798982Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation a9c0f506-a9ea-4a9a-87d2-c2722e9d664e · inbound

From an Image to a Scene: Learning to Imagine the World from a Million 360 Videos cites this paper.

From an Image to a Scene: Learning to Imagine the World from a Million 360 Videos Conditional Object-Centric Learning from Video

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-11T18:33:51.430960Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T18:33:51.430960Z digest=sha256:352bab5747b85359e045fabfb1c03e6f814125eb4a354f28f9c619c30c50b949

Observation c169f45a-0bd8-45e2-bf80-4eac2708b715 · inbound

Efficient Object-centric Representation Learning with Pre-trained Geometric Prior cites this paper.

Efficient Object-centric Representation Learning with Pre-trained Geometric Prior Conditional Object-Centric Learning from Video

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-11T14:14:48.257074Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T14:14:48.257074Z digest=sha256:73bec9defaa88895bd3f7b2c3e8e30b74e455a19080c0ed025759876c0a345c4

Observation 51a92293-b520-4ddb-844b-77e5d36a50db · inbound

Leveraging Color Channel Independence for Improved Unsupervised Object Detection cites this paper.

Leveraging Color Channel Independence for Improved Unsupervised Object Detection Conditional Object-Centric Learning from Video

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-11T11:38:15.935954Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T11:38:15.935954Z digest=sha256:7ac774ce89af1402fa4afc3232bda45f46970702e94a5f3b6cd737a493a2a17c

Observation a0f22dc8-a667-455f-9dd9-f5f4f266f7a3 · inbound

On the Benefits of Instance Decomposition in Video Prediction Models cites this paper.

On the Benefits of Instance Decomposition in Video Prediction Models Conditional Object-Centric Learning from Video

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-10T19:12:00.719119Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T19:12:00.719119Z digest=sha256:a75226e4a31be868d14dfe3af2c6a5135a99b18157f5925da5a854fe77428949

Observation 9add9272-76ee-4003-ab20-5edf72da4010 · inbound

Dreamweaver: Learning Compositional World Models from Pixels cites this paper.

Dreamweaver: Learning Compositional World Models from Pixels Conditional Object-Centric Learning from Video

Reference 2014

Resolution
unresolved
no resolver link, observed 2026-08-10T15:23:54.937112Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T15:23:54.937112Z digest=sha256:92330fded9aac253e0ac09f537d79eb87ee7fa24f62560c291303b9993874158

Observation 46add00e-7cdc-40bd-842a-6116bb1c4f8e · inbound

Identifiable Object Representations under Spatial Ambiguities cites this paper.

Identifiable Object Representations under Spatial Ambiguities Conditional Object-Centric Learning from Video

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-07T05:32:52.782365Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:32:52.782365Z digest=sha256:716abd7a42c79470bcef13073645c20fe9b2a6e81e65913525384909fdae63e9

Observation 0bd2304f-8036-4784-816e-704ec86f90c1 · inbound

Is an object-centric representation beneficial for robotic manipulation ? cites this paper.

Is an object-centric representation beneficial for robotic manipulation ? Conditional Object-Centric Learning from Video

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-06T23:11:50.019492Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:11:50.019492Z digest=sha256:d010676c1cf3f0578b14280fc2b6ccc04908f9d30414d940304c6dbc957a6db8

Observation 1d1ea8cd-8ae3-4943-9981-d4c88f14cd1a · inbound

Dyn-O: Building Structured World Models with Object-Centric Representations cites this paper.

Dyn-O: Building Structured World Models with Object-Centric Representations Conditional Object-Centric Learning from Video

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-06T20:19:04.792172Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:19:04.792172Z digest=sha256:93b8487e4974283bb6b12e82139618b7c39408820ae5b39c2013890646bf0a46

Observation 59372b0a-d681-41f7-b3ac-678cd33d9cb2 · inbound

Discovering and using Spelke segments cites this paper.

Discovering and using Spelke segments Conditional Object-Centric Learning from Video

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-06T15:25:15.139021Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:25:15.139021Z digest=sha256:ba23027a6dfafc1908b5f6699c21f5fd275b5a3325172f699b1d13c9ede94cd7

Observation 0adbfa82-dc8f-4b0a-aedd-5845e68ba72d · inbound

Learning Object-Centric Representations in SAR Images with Multi-Level Feature Fusion cites this paper.

Learning Object-Centric Representations in SAR Images with Multi-Level Feature Fusion Conditional Object-Centric Learning from Video

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-04T19:24:39.845000Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T19:24:39.845000Z digest=sha256:58709260ed15bc79b64d43a0fd925e0dc40e538b2ddfde721dd84ac56344191e

Observation fa4af0f3-2d81-462b-9cf8-cd2d42f8805c · inbound

Spotlighting Task-Relevant Features: Object-Centric Representations for Better Generalization in Robotic Manipulation cites this paper.

Spotlighting Task-Relevant Features: Object-Centric Representations for Better Generalization in Robotic Manipulation Conditional Object-Centric Learning from Video

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-03T07:04:44.994732Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T07:04:44.994732Z digest=sha256:f1e105cb4aacae7a7ee23e3f7ff54b374f8b6944a86e873e1bc54f3f8cd86661

Observation 8d096808-926b-48ce-aaa3-caf2790043bc · inbound

Factored Latent Action World Models cites this paper.

Factored Latent Action World Models Conditional Object-Centric Learning from Video

Reference 2013

Resolution
unresolved
no resolver link, observed 2026-08-02T22:41:09.572161Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T22:41:09.572161Z digest=sha256:03ffff0f3f47da83ea3e4276aa38d374d21f93aa8d31136905047798fc04cb7d

Observation 9a2cd549-db14-4483-b8b2-93cc3839be98 · inbound

ES-Merging: Biological MLLM Merging via Embedding Space Signals cites this paper.

ES-Merging: Biological MLLM Merging via Embedding Space Signals Conditional Object-Centric Learning from Video

Reference 22

Resolution
unresolved
no resolver link, observed 2026-07-14T21:18:21.658797Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T21:18:21.658797Z digest=sha256:8e9f61316c3016f2b9a729d68d2b5c216af1efcb5d154cc9a8b5ae6a6982be0a

Observation d721afb8-31a2-4ad6-a46c-bff25ae7ab7c · inbound

Scene-Agnostic Object-Centric Representation Learning for 3D Gaussian Splatting cites this paper.

Scene-Agnostic Object-Centric Representation Learning for 3D Gaussian Splatting Conditional Object-Centric Learning from Video

Reference 20

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T06:41:33.148395Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-05-10T17:29:54.777890Z digest=sha256:e4cc270e4e767d066751966c1fab51fba7b6ee50f7f9ee2371c4e756bc789b20

Observation e040cfac-6f47-4570-bdee-0cfd7d33dcb7 · inbound

OFlow: Injecting Object-Aware Temporal Flow Matching for Robust Robotic Manipulation cites this paper.

OFlow: Injecting Object-Aware Temporal Flow Matching for Robust Robotic Manipulation Conditional Object-Centric Learning from Video

Reference 32

Resolution
verified exact
arxiv_id, observed 2026-05-10T11:30:20.154986Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-05-10T04:50:33.134927Z digest=sha256:b47bc546741e6a1ce8bf5a5734b654f910bfc20e1bffa67e849f0ca3faebc1c7

Observation 66ec7806-7122-4851-86b2-3b3d5f12f504 · inbound

Unsupervised Learning of Inter-Object Relationships via Group Homomorphism cites this paper.

Unsupervised Learning of Inter-Object Relationships via Group Homomorphism Conditional Object-Centric Learning from Video

Reference 9

Resolution
metadata mismatch
arxiv_id, observed 2026-05-10T00:59:49.777721Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-05-10T00:55:21.788068Z digest=sha256:56ad9921a1cc39eb2ad2a0f3e43a84696272c14c6b1a72221d305d1282694608

Observation 69441e14-2995-4cb1-a443-b0448de741b7 · inbound

OA-WAM: Object-Addressable World Action Model for Robust Robot Manipulation cites this paper.

OA-WAM: Object-Addressable World Action Model for Robust Robot Manipulation Conditional Object-Centric Learning from Video

Reference 35

Resolution
verified exact
arxiv_id, observed 2026-05-11T20:31:11.286707Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-05-08T08:52:07.545191Z digest=sha256:175ceca3efe952c57f8337c71246fca49f676389a66a8f7f08ec526a4fbfa4d8

Observation b4a56c2b-066a-4367-a1fa-017e1e2b793d · inbound

InfoGeo: Information-Theoretic Object-Centric Learning for Cross-View Generalizable UAV Geo-Localization cites this paper.

InfoGeo: Information-Theoretic Object-Centric Learning for Cross-View Generalizable UAV Geo-Localization Conditional Object-Centric Learning from Video

Reference 15

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T04:25:55.694262Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=arxiv_source observed=2026-05-11T01:27:43.340357Z digest=sha256:791bb0cdbdfb626e21c378f15354b351ccdba34bfe83b475855889b005726e95

Observation b8085708-8806-4ae1-a1e9-c0a623cd1b25 · inbound

Dual-State Slot Attention: Decoupling Appearance and Identity for Video Object-Centric Learning cites this paper.

Dual-State Slot Attention: Decoupling Appearance and Identity for Video Object-Centric Learning Conditional Object-Centric Learning from Video

Reference 7

Resolution
metadata mismatch
arxiv_id, observed 2026-07-03T10:48:02.800565Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=arxiv_source observed=2026-06-27T09:49:52.527042Z digest=sha256:2390ec93410505478cc91baea27098ced1bd79a58a3fff82b853ae556da8fed1