Pith. sign in

Paper Citation Record · LEDGER

QUART-Online: Latency-Free Large Multimodal Language Model for Quadruped Robot Learning

As of 15 August 2026, this Paper Citation Record lists 41 of 41 outbound references and 1 inbound Pith citation observation for arXiv:2412.15576.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2412.15576 v5

Coverage vector

measured 41 of 41 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-11T11:22:41.433077Z

measured 42 of 42 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-15T06:32:42.880941+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-05T20:32:46.642262Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-08-05T20:32:48.204027Z

Reference resolution

41 of 41 outbound references displayed

  • verified exact1
  • verified fuzzy17
  • unresolved23
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 60be2f5f-79f3-47b0-9633-a3a14e98fc58 · outbound

This paper cites SayTap: Language to Quadrupedal Locomotion.

QUART-Online: Latency-Free Large Multimodal Language Model for Quadruped Robot Learning SayTap: Language to Quadrupedal Locomotion

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-11T11:22:41.217166Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T11:22:41.217166Z digest=sha256:2811e28f33a94505de3dc6956a3d359ecb4c9a04236fe2a85a5bff53732d212f

Observation 74a0ee8e-f56b-4c45-841a-ae3e4a946001 · outbound

This paper cites Vinl: Visual navigation and locomotion over obstacles,.

QUART-Online: Latency-Free Large Multimodal Language Model for Quadruped Robot Learning Vinl: Visual navigation and locomotion over obstacles,

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T11:22:42.197817Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-11T11:22:41.223887Z digest=sha256:f72c25a9d3739109436d49fd80284b5d66adcd6bc69f776f6025e721e5c4d93b

Observation 645a456c-9e9c-4d1b-a674-dc8d90c72b1f · outbound

This paper cites Socially compliant navigation dataset (scand): A large-scale dataset of demonstrations for social navigation,.

QUART-Online: Latency-Free Large Multimodal Language Model for Quadruped Robot Learning Socially compliant navigation dataset (scand): A large-scale dataset of demonstrations for social navigation,

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T11:22:42.177244Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-11T11:22:41.229309Z digest=sha256:3aa11631ee797dfd7680e5edf0fee6b7e6bc6a0ef05904751ebb7c12f1367b50

Observation 57d4e880-6d48-4f55-b3ea-6ba9a1f4a8fb · outbound

This paper cites QUAR-VLA: Vision-Language-Action Model for Quadruped Robots.

QUART-Online: Latency-Free Large Multimodal Language Model for Quadruped Robot Learning QUAR-VLA: Vision-Language-Action Model for Quadruped Robots

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-11T11:22:41.234636Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T11:22:41.234636Z digest=sha256:0b058c404253908158f3647efdb644e4781ebfbf112fcf235f424af847e5b7ee

Observation aa14e0f0-eeed-467c-b059-e0665fc2265a · outbound

This paper cites Barkour: Benchmarking Animal-level Agility with Quadruped Robots.

QUART-Online: Latency-Free Large Multimodal Language Model for Quadruped Robot Learning Barkour: Benchmarking Animal-level Agility with Quadruped Robots

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-11T11:22:41.240663Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T11:22:41.240663Z digest=sha256:08adbcc3573dcb9f64e025a0495fa897c2f388f11ec9e62ebb012d82ef6b6565

Observation e84992c1-0115-4554-8fb7-3f5fc9a28cbe · outbound

This paper cites Learning Vision-Guided Quadrupedal Locomotion End-to-End with Cross-Modal Transformers.

QUART-Online: Latency-Free Large Multimodal Language Model for Quadruped Robot Learning Learning Vision-Guided Quadrupedal Locomotion End-to-End with Cross-Modal Transformers

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-11T11:22:41.246078Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T11:22:41.246078Z digest=sha256:185dd8f52a07de715fa1c2734478479837d8cdfc97df5faa3e48713c4f203270

Observation 2d2f25df-0ca2-4790-ba4b-080f0775c99e · outbound

This paper cites Learning advanced locomotion for quadrupedal robots: A distributed multi-agent reinforcement learning framework with riemannian motion policies,.

QUART-Online: Latency-Free Large Multimodal Language Model for Quadruped Robot Learning Learning advanced locomotion for quadrupedal robots: A distributed multi-agent reinforcement learning framework with riemannian motion policies,

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T11:22:42.157961Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-11T11:22:41.252377Z digest=sha256:a07f6fc1d975bc37c0e0f321da8625e093f6a9aa81668c41124b2287aacc8ae5

Observation 9119a7a6-b919-4045-83fb-b1b08414bf85 · outbound

This paper cites The surprising effectiveness of representation learning for visual imitation,.

QUART-Online: Latency-Free Large Multimodal Language Model for Quadruped Robot Learning The surprising effectiveness of representation learning for visual imitation,

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T11:22:42.135457Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-11T11:22:41.257467Z digest=sha256:332c87fbfe064b40646b6cc5298550ca69873a88a273aae471dc4ed0e67f6cb2

Observation d6111d79-c46a-4990-beb5-2d41fbf48853 · outbound

This paper cites Learning agent-aware affordances for closed-loop interaction with articulated objects,.

QUART-Online: Latency-Free Large Multimodal Language Model for Quadruped Robot Learning Learning agent-aware affordances for closed-loop interaction with articulated objects,

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T11:22:42.111427Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-11T11:22:41.262087Z digest=sha256:668607045c816402e7e6c0f932f7f359cf8ddcdf13b4be711f1783f3109104c7

Observation 2d57eeaa-2be0-49a3-8034-16cc6767e686 · outbound

This paper cites Affordances from human videos as a versatile representation for robotics,.

QUART-Online: Latency-Free Large Multimodal Language Model for Quadruped Robot Learning Affordances from human videos as a versatile representation for robotics,

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T11:22:42.090893Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-11T11:22:41.268521Z digest=sha256:3bc9bba1b2189dd4120b26ab916b6b079f88e5fac5343aaa7fecea4549031be9

Observation 95cc8283-4ad9-4f65-98bc-87b4a70c87d9 · outbound

This paper cites Making Sense of Vision and Touch: Self-Supervised Learning of Multimodal Representations for Contact-Rich Tasks.

QUART-Online: Latency-Free Large Multimodal Language Model for Quadruped Robot Learning Making Sense of Vision and Touch: Self-Supervised Learning of Multimodal Representations for Contact-Rich Tasks

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-11T11:22:41.273478Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T11:22:41.273478Z digest=sha256:2ace1835481dddf9dcd06b1b78f86ae00727d8cf894e5736e4513577ac8118bd

Observation e93bf7f0-d584-4f54-a03e-5e172e3453af · outbound

This paper cites Hydra: Hybrid robot actions for imitation learning,.

QUART-Online: Latency-Free Large Multimodal Language Model for Quadruped Robot Learning Hydra: Hybrid robot actions for imitation learning,

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T11:22:42.072925Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-11T11:22:41.278491Z digest=sha256:f656b64da69a9e55ea2428990bae2e52798f3c8a1be546cf0add0040fdf43b47

Observation 72c907f7-ea3f-474b-b2cb-992e5aee9dc4 · outbound

This paper cites Watch and match: Supercharging imitation with regularized optimal transport,.

QUART-Online: Latency-Free Large Multimodal Language Model for Quadruped Robot Learning Watch and match: Supercharging imitation with regularized optimal transport,

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T11:22:42.054274Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-11T11:22:41.282971Z digest=sha256:74242381806b75eadb0f907cd745c97f5f8a091de45a51e57c7c6bf9a7d65c4b

Observation 73362a02-a32c-4394-b82d-4e7cdd2a4ac9 · outbound

This paper cites Learning robust perceptive locomotion for quadrupedal robots in the wild,.

QUART-Online: Latency-Free Large Multimodal Language Model for Quadruped Robot Learning Learning robust perceptive locomotion for quadrupedal robots in the wild,

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-11T11:22:41.287533Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T11:22:41.287533Z digest=sha256:d71da0d9b8498c27be6b90fac91fd84966f896ebd501b330ba9a5534cae6ce90

Observation 01285e98-4e15-4ba9-8908-759e17922034 · outbound

This paper cites Learning Visual Quadrupedal Loco-Manipulation from Demonstrations.

QUART-Online: Latency-Free Large Multimodal Language Model for Quadruped Robot Learning Learning Visual Quadrupedal Loco-Manipulation from Demonstrations

Reference 15

Resolution
verified exact
local_arxiv, observed 2026-08-11T11:22:41.735895Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-11T11:22:41.292209Z digest=sha256:c1137551f523cf403962fd8008b6ac7630ee6d3a6f41370c114fec183f135b96

Observation 794d358f-67c4-4e8b-ac42-8cce1771588f · outbound

This paper cites GenLoco: Generalized Locomotion Controllers for Quadrupedal Robots.

QUART-Online: Latency-Free Large Multimodal Language Model for Quadruped Robot Learning GenLoco: Generalized Locomotion Controllers for Quadrupedal Robots

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-11T11:22:41.297592Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T11:22:41.297592Z digest=sha256:a669421224aed2f5afb2e7e84fcb6c9ba4f69ecfa84354cb1cb9aaa27b5139a1

Observation 1b4ca266-b4b6-47c2-a2fe-85663f10739f · outbound

This paper cites Learning whole-body manipulation for quadrupedal robot,.

QUART-Online: Latency-Free Large Multimodal Language Model for Quadruped Robot Learning Learning whole-body manipulation for quadrupedal robot,

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-11T11:22:41.302510Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T11:22:41.302510Z digest=sha256:6126716a34614183e07b3bad83178987052409cca051833cd5de972e813ea107

Observation 4328afb8-011c-44b7-953b-09bb33b1fdaa · outbound

This paper cites Long-horizon Locomotion and Manipulation on a Quadrupedal Robot with Large Language Models.

QUART-Online: Latency-Free Large Multimodal Language Model for Quadruped Robot Learning Long-horizon Locomotion and Manipulation on a Quadrupedal Robot with Large Language Models

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-11T11:22:41.307040Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T11:22:41.307040Z digest=sha256:a206a52b739a8429d2683f6935333d4962356ae004a92140a283ce8e46cdce13

Observation 401ede30-e458-4817-9c90-0684d036018e · outbound

This paper cites Lanmp: A multifaceted mobile manipulation benchmark for robots,.

QUART-Online: Latency-Free Large Multimodal Language Model for Quadruped Robot Learning Lanmp: A multifaceted mobile manipulation benchmark for robots,

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T11:22:42.006338Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-11T11:22:41.313846Z digest=sha256:7dfdbb4acbdd89d349c564357f60c8bb00436ff9531d9268b51f0f912ef93e52

Observation b9c07f87-13ca-40d2-8305-fb4974828e23 · outbound

This paper cites Learning Multiple Gaits within Latent Space for Quadruped Robots.

QUART-Online: Latency-Free Large Multimodal Language Model for Quadruped Robot Learning Learning Multiple Gaits within Latent Space for Quadruped Robots

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-11T11:22:41.318950Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T11:22:41.318950Z digest=sha256:8afd891eec1a1b6ee7ec51c6e3df557d0d75ab9fdf6315cf0775955faa29246d

Observation a5e12615-f93f-4374-9890-2c03e51f3be6 · outbound

This paper cites QuadrupedGPT: Towards a Versatile Quadruped Agent in Open-ended Worlds.

QUART-Online: Latency-Free Large Multimodal Language Model for Quadruped Robot Learning QuadrupedGPT: Towards a Versatile Quadruped Agent in Open-ended Worlds

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-11T11:22:41.324250Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T11:22:41.324250Z digest=sha256:6ac990b2b0d7fe1967d00dc307fa1299389619f154965f767cbb425c482812a8

Observation 80b60522-054a-475e-8f13-61ccd396563b · outbound

This paper cites Germ: A generalist robotic model with mixture-of-experts for quadruped robot,.

QUART-Online: Latency-Free Large Multimodal Language Model for Quadruped Robot Learning Germ: A generalist robotic model with mixture-of-experts for quadruped robot,

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T11:22:41.984765Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-11T11:22:41.329492Z digest=sha256:23ce225eb4f11de4c2860230103372fa532329636dcf09c880f2a0a768080bc0

Observation 669082bf-1ecf-42f6-8c71-412ab3137cb0 · outbound

This paper cites Efficient Large Language Models: A Survey.

QUART-Online: Latency-Free Large Multimodal Language Model for Quadruped Robot Learning Efficient Large Language Models: A Survey

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-11T11:22:41.334325Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T11:22:41.334325Z digest=sha256:02f72fc2b6e86afc7ea80e7fa81ad935b37fa032dca5be7be3e5f9b646780d06

Observation f53b2776-5a3e-4ef6-b1bf-cab985476144 · outbound

This paper cites Dynamic neural networks: A survey,.

QUART-Online: Latency-Free Large Multimodal Language Model for Quadruped Robot Learning Dynamic neural networks: A survey,

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T11:22:41.967075Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-11T11:22:41.340281Z digest=sha256:79e4b8812a16406665772686688e44836f0b9e6cdcd215b022b2ed0c536356c6

Observation 395eb0f1-2869-411e-88a3-1acf333e97fc · outbound

This paper cites LLaVA-Gemma: Accelerating Multimodal Foundation Models with a Compact Language Model.

QUART-Online: Latency-Free Large Multimodal Language Model for Quadruped Robot Learning LLaVA-Gemma: Accelerating Multimodal Foundation Models with a Compact Language Model

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-11T11:22:41.346370Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T11:22:41.346370Z digest=sha256:43600c906b4bcd7a6333a22692035878dd8e7d18502b9933b7f1edd95b2bd13e

Observation 74887ac0-9896-4195-a910-840ceb9d3af4 · outbound

This paper cites MoE-LLaVA: Mixture of Experts for Large Vision-Language Models.

QUART-Online: Latency-Free Large Multimodal Language Model for Quadruped Robot Learning MoE-LLaVA: Mixture of Experts for Large Vision-Language Models

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-11T11:22:41.352432Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T11:22:41.352432Z digest=sha256:087cad3191a33c559754d24b4b7932c2503fe2cf71f9602c9b3a1576396c2acb

Observation d598089a-bbb6-46fa-a662-b988adb0a397 · outbound

This paper cites Cobra: Extending Mamba to Multi-Modal Large Language Model for Efficient Inference.

QUART-Online: Latency-Free Large Multimodal Language Model for Quadruped Robot Learning Cobra: Extending Mamba to Multi-Modal Large Language Model for Efficient Inference

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-11T11:22:41.358407Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T11:22:41.358407Z digest=sha256:30bc3e05e1ef681db6b1981601d6027ab1810fa90f0e103a072ec9b9a342a731

Observation 132453f2-c595-4cf9-bf01-c29543811d4e · outbound

This paper cites Linearizing Large Language Models.

QUART-Online: Latency-Free Large Multimodal Language Model for Quadruped Robot Learning Linearizing Large Language Models

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-11T11:22:41.363425Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T11:22:41.363425Z digest=sha256:2f96ef622832681a8f98d47cd78fb179325e4c522d824555337858bc4a98c13a

Observation bca453a0-1246-47ac-bcce-01962c9c09f8 · outbound

This paper cites Rt-1: Robotics transformer for real-world control at scale,.

QUART-Online: Latency-Free Large Multimodal Language Model for Quadruped Robot Learning Rt-1: Robotics transformer for real-world control at scale,

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T11:22:41.950244Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-11T11:22:41.368805Z digest=sha256:e36358a6010417f86c349d6283ccc2aef4f20a3d82b45751fa4f0974687c37e3

Observation 5c92a805-34b5-4ba6-931b-bbd95b3ddbf4 · outbound

This paper cites Rt-2: Vision-language-action models transfer web knowledge to robotic control,.

QUART-Online: Latency-Free Large Multimodal Language Model for Quadruped Robot Learning Rt-2: Vision-language-action models transfer web knowledge to robotic control,

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T11:22:41.931148Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-11T11:22:41.373603Z digest=sha256:462a3c8c1f89afac456e1a81a115281c4bae3382850ec45314878a685f565707

Observation 2dc9164e-718a-4ad9-9eeb-2a683d1b223e · outbound

This paper cites RT-H: Action Hierarchies Using Language.

QUART-Online: Latency-Free Large Multimodal Language Model for Quadruped Robot Learning RT-H: Action Hierarchies Using Language

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-11T11:22:41.378445Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T11:22:41.378445Z digest=sha256:5bb77053791d64bdfe16bb1dec67924783abde07963b49ee1c150d09e588efa3

Observation 8fac57c2-e007-4d3b-9964-2d1106c9d113 · outbound

This paper cites OpenVLA: An Open-Source Vision-Language-Action Model.

QUART-Online: Latency-Free Large Multimodal Language Model for Quadruped Robot Learning OpenVLA: An Open-Source Vision-Language-Action Model

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-11T11:22:41.383200Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T11:22:41.383200Z digest=sha256:15ded81a958f547a61f97fbb908f15932ddb74047d5445da3f593acc310a9181

Observation 4a3b9cf2-45f3-4e54-a39d-fb8ac3bc1386 · outbound

This paper cites Unleashing Large-Scale Video Generative Pre-training for Visual Robot Manipulation.

QUART-Online: Latency-Free Large Multimodal Language Model for Quadruped Robot Learning Unleashing Large-Scale Video Generative Pre-training for Visual Robot Manipulation

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-11T11:22:41.388623Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T11:22:41.388623Z digest=sha256:e78721d82d57dd5c59c0318ab31a4f832b94b9e2b2ce1057a059beadc14c63a5

Observation b620cb81-1c65-4201-ac9c-6e1e5feead9c · outbound

This paper cites Vision-Language Foundation Models as Effective Robot Imitators.

QUART-Online: Latency-Free Large Multimodal Language Model for Quadruped Robot Learning Vision-Language Foundation Models as Effective Robot Imitators

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-11T11:22:41.393711Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T11:22:41.393711Z digest=sha256:65f4b79bc6ac04f2c11a314f56849a56b1d4a0af2798791df59d02b265491ffd

Observation be9801a6-3dd5-4500-af28-034933a14083 · outbound

This paper cites 3D Diffuser Actor: Policy Diffusion with 3D Scene Representations.

QUART-Online: Latency-Free Large Multimodal Language Model for Quadruped Robot Learning 3D Diffuser Actor: Policy Diffusion with 3D Scene Representations

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-11T11:22:41.400080Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T11:22:41.400080Z digest=sha256:8c9e340e9f83c7176fb9e7ef70436baeffe0cfaeb5de3c683e1ff7f4b445ce26

Observation 2608c188-55e2-4c61-b7e3-fa4c413ad1dd · outbound

This paper cites Grounding Multimodal Large Language Models in Actions.

QUART-Online: Latency-Free Large Multimodal Language Model for Quadruped Robot Learning Grounding Multimodal Large Language Models in Actions

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-11T11:22:41.405361Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T11:22:41.405361Z digest=sha256:5d6ec10dfd042b47f6a02e8214544216fd403098bf9cb934bd6b1667104a54d5

Observation 420716c5-83a6-4d25-9df9-c802eb88c732 · outbound

This paper cites Introducing our multimodal models,.

QUART-Online: Latency-Free Large Multimodal Language Model for Quadruped Robot Learning Introducing our multimodal models,

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T11:22:41.911325Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-11T11:22:41.410576Z digest=sha256:cd4d2423d29c9bc6c2b21bd54f342ea1415d8efaf280eb859d8ecc034f6a8701

Observation 13485335-b93e-4c28-95e2-33307a4a404b · outbound

This paper cites Isaac gym: High performance gpu-based physics simulation for robot learning,.

QUART-Online: Latency-Free Large Multimodal Language Model for Quadruped Robot Learning Isaac gym: High performance gpu-based physics simulation for robot learning,

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-11T11:22:41.415810Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T11:22:41.415810Z digest=sha256:61709a4449a6fabea0514ae38504c53925e6f055777117d77eaa84570eaab6cd

Observation 3b4d8d08-a0c0-4863-bd73-329914abb50f · outbound

This paper cites Walk these ways: Tuning robot control for generalization with multiplicity of behavior,.

QUART-Online: Latency-Free Large Multimodal Language Model for Quadruped Robot Learning Walk these ways: Tuning robot control for generalization with multiplicity of behavior,

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T11:22:41.879852Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-11T11:22:41.421612Z digest=sha256:dd8ee7dc321199d49f2c84a79dcdf57fe451d9bbec8cac978db8f28f66a30a1f

Observation 06a5f743-64ed-4910-9b0a-147f25714b4d · outbound

This paper cites Learning transferable visual models from natural language supervision,.

QUART-Online: Latency-Free Large Multimodal Language Model for Quadruped Robot Learning Learning transferable visual models from natural language supervision,

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T11:22:41.863128Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-11T11:22:41.426831Z digest=sha256:5f432598b877bf69eba7e817dcc3eb280a8b74d3117a34b7b637016dafd9ed90

Observation 3111d6e9-57e9-47d3-80f1-4444d28df180 · outbound

This paper cites Where are we in the search for an artificial visual cortex for embodied intelligence?.

QUART-Online: Latency-Free Large Multimodal Language Model for Quadruped Robot Learning Where are we in the search for an artificial visual cortex for embodied intelligence?

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T11:22:41.845090Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-11T11:22:41.433077Z digest=sha256:b5be4e6d81c93a15259dea29e5d06bebfda54358bab2ee51bc261549fbc1dbd5

Pith citing papers

Observation 04fd7e57-eaf3-4ceb-b882-4eea9e122eba · inbound

ReconVLA: Reconstructive Vision-Language-Action Model as Effective Robot Perceiver cites this paper.

ReconVLA: Reconstructive Vision-Language-Action Model as Effective Robot Perceiver QUART-Online: Latency-Free Large Multimodal Language Model for Quadruped Robot Learning

Reference 45

Resolution
verified exact
local_arxiv, observed 2026-08-05T20:32:48.264716Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=arxiv_source observed=2026-08-05T20:32:46.642262Z digest=sha256:90774a92fa44322abb50fffae0253f4ed2929522f980d1e596b9bc54e9c194aa