Pith. sign in

Paper Citation Record · LEDGER

MASH: Cooperative-Heterogeneous Multi-Agent Reinforcement Learning for Single Humanoid Robot Locomotion

As of 18 August 2026, this Paper Citation Record lists 47 of 47 outbound references and 0 inbound Pith citation observations for arXiv:2508.10423.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2508.10423 v1

Coverage vector

measured 47 of 47 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-05T20:30:30.203171Z

measured 47 of 47 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-18T06:34:40.430872+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

47 of 47 outbound references displayed

  • verified exact2
  • verified fuzzy37
  • unresolved8
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 77978675-00ca-4485-8090-d5e4154448f3 · outbound

This paper cites HOVER: Versatile neural whole-body controller for humanoid robots,.

MASH: Cooperative-Heterogeneous Multi-Agent Reinforcement Learning for Single Humanoid Robot Locomotion HOVER: Versatile neural whole-body controller for humanoid robots,

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:30:31.115929Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-05T20:30:26.580920Z digest=sha256:bdc9f01f68c7f7f2fd395c0fad739d4256e6da993c7e2ae9454ad101bd7f8bc0

Observation a7977c1f-a96e-4aa1-8d43-42653518fb8a · outbound

This paper cites Distributional policy gradient with distributional value function,.

MASH: Cooperative-Heterogeneous Multi-Agent Reinforcement Learning for Single Humanoid Robot Locomotion Distributional policy gradient with distributional value function,

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:30:31.098429Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-05T20:30:26.649591Z digest=sha256:13554a9cdd156c30e76f4fe18d1632cbc192240bd04b95fef3c85e9dbf870e76

Observation fe4023d4-eb8b-4e26-a495-6bf5f45352f6 · outbound

This paper cites Walk these ways: Tuning robot control for generalization with multiplicity of behavior,.

MASH: Cooperative-Heterogeneous Multi-Agent Reinforcement Learning for Single Humanoid Robot Locomotion Walk these ways: Tuning robot control for generalization with multiplicity of behavior,

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:30:31.081654Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-05T20:30:26.738893Z digest=sha256:0949118ad3f1c4e90b372f9913cfa959f725b4cddb934637a10310bb5dabe0b1

Observation 985e80ca-683d-4f4e-bf17-4fa640ee2d8b · outbound

This paper cites Dreamwaq: Learning robust quadrupedal locomotion with implicit terrain imagination via deep reinforcement learning,.

MASH: Cooperative-Heterogeneous Multi-Agent Reinforcement Learning for Single Humanoid Robot Locomotion Dreamwaq: Learning robust quadrupedal locomotion with implicit terrain imagination via deep reinforcement learning,

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:30:31.063656Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-05T20:30:26.795999Z digest=sha256:8ca5c1abe4c832961bc5732a8b1c332de28bb77404bf312615f5e25b1ec57b8c

Observation 7c9e068d-c598-4426-a027-ebc4b4664807 · outbound

This paper cites Biped dynamic walking using reinforcement learning,.

MASH: Cooperative-Heterogeneous Multi-Agent Reinforcement Learning for Single Humanoid Robot Locomotion Biped dynamic walking using reinforcement learning,

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:30:31.045484Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-05T20:30:26.847516Z digest=sha256:ff6adcd211e76ef1bb909b028569f9967b8e3eb29805dae4ed60ce68c3e0803a

Observation c396c3dd-413b-41e3-8349-abb58f1d29e6 · outbound

This paper cites Learning vision-based bipedal locomotion for challeng- ing terrain,.

MASH: Cooperative-Heterogeneous Multi-Agent Reinforcement Learning for Single Humanoid Robot Locomotion Learning vision-based bipedal locomotion for challeng- ing terrain,

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:30:31.024478Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-05T20:30:26.945183Z digest=sha256:7f9898c10bb557a2d65586cb76d18e186f6b766f68294f78ce0fe0b05b434ec2

Observation 0ee0e265-40be-47bb-83e0-2ab770211a81 · outbound

This paper cites Motor anomaly detection for unmanned aerial vehicles using reinforcement learning,.

MASH: Cooperative-Heterogeneous Multi-Agent Reinforcement Learning for Single Humanoid Robot Locomotion Motor anomaly detection for unmanned aerial vehicles using reinforcement learning,

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:30:30.993595Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-05T20:30:27.019704Z digest=sha256:2ed54231521bc8831fc0067bdf6e5f1c7c1d66346c78b9c444ea7f4969beaf97

Observation c919e3cd-cf6a-4840-a824-a510f9f64f00 · outbound

This paper cites Optimization-based control for dynamic legged robots,.

MASH: Cooperative-Heterogeneous Multi-Agent Reinforcement Learning for Single Humanoid Robot Locomotion Optimization-based control for dynamic legged robots,

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:30:30.972502Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-05T20:30:27.111040Z digest=sha256:6c52421d9113968d9c58dac8a417f334c223ddcbdc84f0945dddd80d0bd29787

Observation 6b8305bd-9467-4bbb-a3b7-aca9f7e45c64 · outbound

This paper cites Versatile multicontact planning and control for legged loco-manipulation,.

MASH: Cooperative-Heterogeneous Multi-Agent Reinforcement Learning for Single Humanoid Robot Locomotion Versatile multicontact planning and control for legged loco-manipulation,

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-05T20:30:27.279483Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T20:30:27.279483Z digest=sha256:c94526575e3026c50b51efa479fc399451d8dbb34e6b77d2acf068da36e6dbb6

Observation 34a470d6-66e6-426a-89b5-32f07bea1041 · outbound

This paper cites Combining trajectory optimization, supervised machine learning, and model structure for mitigating the curse of dimensionality in the control of bipedal robots,.

MASH: Cooperative-Heterogeneous Multi-Agent Reinforcement Learning for Single Humanoid Robot Locomotion Combining trajectory optimization, supervised machine learning, and model structure for mitigating the curse of dimensionality in the control of bipedal robots,

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-05T20:30:27.398322Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T20:30:27.398322Z digest=sha256:3c1ac70f13483f5199d3e2ad6bcd23fee36b1fecbf6270c947ab8681e5cf413c

Observation db20e7da-e494-42b6-9fa5-c72d95d9758c · outbound

This paper cites Real-world humanoid locomotion with reinforcement learning,.

MASH: Cooperative-Heterogeneous Multi-Agent Reinforcement Learning for Single Humanoid Robot Locomotion Real-world humanoid locomotion with reinforcement learning,

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-05T20:30:27.567294Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T20:30:27.567294Z digest=sha256:b56443ab4781cab7558857c665d82c7c6a33d47192ac591051858abbc2e72abd

Observation 796fbf8d-98bb-459e-9f76-c6f243150cdc · outbound

This paper cites Not only rewards but also constraints: Applications on legged robot locomotion,.

MASH: Cooperative-Heterogeneous Multi-Agent Reinforcement Learning for Single Humanoid Robot Locomotion Not only rewards but also constraints: Applications on legged robot locomotion,

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-05T20:30:27.661298Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T20:30:27.661298Z digest=sha256:63c3b4fe1a7f1d6bb150951c48ca8f1ede56bd567655bd3400da8d6edce39881

Observation 4c7eb530-016d-42ac-9dea-1a50c79ae653 · outbound

This paper cites Diffusion policy: Visuomotor policy learning via ac- tion diffusion,.

MASH: Cooperative-Heterogeneous Multi-Agent Reinforcement Learning for Single Humanoid Robot Locomotion Diffusion policy: Visuomotor policy learning via ac- tion diffusion,

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-05T20:30:27.776297Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T20:30:27.776297Z digest=sha256:25c02f0819ec7546df1cf977360aa17ef6e223b53161da04e824ea9112ebe8c3

Observation deca8a10-e073-4628-953f-b38508e1f524 · outbound

This paper cites Learning-based legged locomotion: State of the art and future perspectives,.

MASH: Cooperative-Heterogeneous Multi-Agent Reinforcement Learning for Single Humanoid Robot Locomotion Learning-based legged locomotion: State of the art and future perspectives,

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:30:30.888448Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-05T20:30:27.906470Z digest=sha256:cf54b5c4ab62d4a7c624bc58ed98692a922287cf4ef39ebbeb82fcc28b8c1742

Observation 4becff69-4421-4263-b090-a7df74190e7a · outbound

This paper cites Learning whole-body loco-manipulation for omni-directional task space pose tracking with a wheeled-quadrupedal-manipulator,.

MASH: Cooperative-Heterogeneous Multi-Agent Reinforcement Learning for Single Humanoid Robot Locomotion Learning whole-body loco-manipulation for omni-directional task space pose tracking with a wheeled-quadrupedal-manipulator,

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:30:30.870338Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-05T20:30:28.024263Z digest=sha256:3f84366de4dd2d1d299f628c1d7d264b861796ae869049698265843789090e80

Observation 46252b04-f802-4a0a-88e7-4d329b2778cf · outbound

This paper cites Learning agile soccer skills for a bipedal robot with deep reinforcement learning,.

MASH: Cooperative-Heterogeneous Multi-Agent Reinforcement Learning for Single Humanoid Robot Locomotion Learning agile soccer skills for a bipedal robot with deep reinforcement learning,

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:30:30.853551Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-05T20:30:28.188458Z digest=sha256:04727c8e5baac33a48845fcbbeb92ad98ddee9699b0d3403c00df5f56f08db87

Observation cd324124-89f6-4c88-95c0-b2eacca536ee · outbound

This paper cites Visual whole-body control for legged loco-manipulation,.

MASH: Cooperative-Heterogeneous Multi-Agent Reinforcement Learning for Single Humanoid Robot Locomotion Visual whole-body control for legged loco-manipulation,

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:30:30.837324Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-05T20:30:28.307712Z digest=sha256:5c4a5cd960c05cc3c53ad5c1304dbf852f7137776e813d640b19322ffad7807a

Observation bbbe3bc7-956b-48c5-b565-3979b034076e · outbound

This paper cites Teleoperation of humanoid robots: A survey,.

MASH: Cooperative-Heterogeneous Multi-Agent Reinforcement Learning for Single Humanoid Robot Locomotion Teleoperation of humanoid robots: A survey,

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:30:30.822545Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-05T20:30:28.432305Z digest=sha256:242f45cad3286cc26f92956fbec762ab022121483d465c67bdd8f6f215630a7f

Observation b80c870b-c2c3-448b-bcea-1eaf90a43018 · outbound

This paper cites Sim-to-real robotic sketching using behavior cloning and reinforcement learning,.

MASH: Cooperative-Heterogeneous Multi-Agent Reinforcement Learning for Single Humanoid Robot Locomotion Sim-to-real robotic sketching using behavior cloning and reinforcement learning,

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:30:30.805259Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-05T20:30:28.565492Z digest=sha256:d7a12adb86b78e9607674dd3cbb7625647bbbb51446848fe4345afff3383b8eb

Observation a823bc1a-a4d1-4025-b640-27f1e7692334 · outbound

This paper cites A composite control strategy for quadruped robot by integrating reinforcement learning and model-based control,.

MASH: Cooperative-Heterogeneous Multi-Agent Reinforcement Learning for Single Humanoid Robot Locomotion A composite control strategy for quadruped robot by integrating reinforcement learning and model-based control,

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:30:30.787160Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-05T20:30:28.728125Z digest=sha256:45485c063e966b6df67d26732446c18334ef253f07c8db77a438ef57bce7dbb5

Observation edfb33f5-fe21-4ca7-b58f-03b78e567cab · outbound

This paper cites Towards human-level bimanual dexterous manipulation with reinforcement learning,.

MASH: Cooperative-Heterogeneous Multi-Agent Reinforcement Learning for Single Humanoid Robot Locomotion Towards human-level bimanual dexterous manipulation with reinforcement learning,

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:30:30.770371Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-05T20:30:28.852754Z digest=sha256:8948a391717b8cdc6cd1af6e8ac9bda861ec4d3aece390575c796299d26ea7ec

Observation 3ea2df64-31e5-4e72-b1f1-1387ac8df8af · outbound

This paper cites Monotonic value function factorisation for deep multi- agent reinforcement learning,.

MASH: Cooperative-Heterogeneous Multi-Agent Reinforcement Learning for Single Humanoid Robot Locomotion Monotonic value function factorisation for deep multi- agent reinforcement learning,

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:30:30.753663Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-05T20:30:29.025150Z digest=sha256:094e1fe85053861ae20fde0e31f555677682ce7769e332ff5c2ff20e052f06e4

Observation 7c4e3bfc-c2bd-4179-8d5b-6abc35ee01bf · outbound

This paper cites Data efficient deep reinforcement learning with action-ranked temporal difference learning,.

MASH: Cooperative-Heterogeneous Multi-Agent Reinforcement Learning for Single Humanoid Robot Locomotion Data efficient deep reinforcement learning with action-ranked temporal difference learning,

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:30:30.736318Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-05T20:30:29.167712Z digest=sha256:871df350490db76acb3109602d2fc75cee4941d968e54f572895a0d6bc42ee77

Observation 578cc9d9-97ad-47af-97ce-8f23987a7746 · outbound

This paper cites MASQ: Multi-Agent Reinforcement Learning for Single Quadruped Robot Locomotion.

MASH: Cooperative-Heterogeneous Multi-Agent Reinforcement Learning for Single Humanoid Robot Locomotion MASQ: Multi-Agent Reinforcement Learning for Single Quadruped Robot Locomotion

Reference 24

Resolution
verified exact
local_arxiv, observed 2026-08-05T20:30:30.324419Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-05T20:30:29.314929Z digest=sha256:a776dde5a66048a11d110817cea1d6974e9fb6b1153005dc2f97cec91d089be1

Observation e95550f3-e507-4b68-b817-530007f7ea10 · outbound

This paper cites Expressive whole-body control for humanoid robots,.

MASH: Cooperative-Heterogeneous Multi-Agent Reinforcement Learning for Single Humanoid Robot Locomotion Expressive whole-body control for humanoid robots,

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:30:30.719256Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-05T20:30:29.389280Z digest=sha256:ce1539f05c1cdc1826ee0419f94ff5d0208a03e56f4f02628b33f71133d660d3

Observation 33c3c267-bcbf-482c-ac3e-13ba11571aa4 · outbound

This paper cites Mobile-television: Predictive motion priors for hu- manoid whole-body control,.

MASH: Cooperative-Heterogeneous Multi-Agent Reinforcement Learning for Single Humanoid Robot Locomotion Mobile-television: Predictive motion priors for hu- manoid whole-body control,

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:30:30.697923Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-05T20:30:29.460025Z digest=sha256:d97e2a51febe6f9abfeb71a79054546f81d60ca1a6e9db5bf3366806f8cb66ba

Observation c211d33c-216c-4a26-977a-e24b062e3fac · outbound

This paper cites Wococo: Learning whole-body humanoid control with sequential contacts,.

MASH: Cooperative-Heterogeneous Multi-Agent Reinforcement Learning for Single Humanoid Robot Locomotion Wococo: Learning whole-body humanoid control with sequential contacts,

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:30:30.677407Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-05T20:30:29.550583Z digest=sha256:60feac68323dc365721c1752903054787c79707184455f56cd645225868b8268

Observation 58e79e0c-8b28-4c1c-a432-294f41703d36 · outbound

This paper cites Learning human-to-humanoid real-time whole-body teleoperation,.

MASH: Cooperative-Heterogeneous Multi-Agent Reinforcement Learning for Single Humanoid Robot Locomotion Learning human-to-humanoid real-time whole-body teleoperation,

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:30:30.659984Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-05T20:30:29.637244Z digest=sha256:f726ddea47ed1cf9ff5e9df884ac050c3434de3fd58f323b4c941417245e8da2

Observation a15a4892-8a69-4a42-8efe-d4763eda588c · outbound

This paper cites Humanplus: Humanoid shadowing and imitation from humans,.

MASH: Cooperative-Heterogeneous Multi-Agent Reinforcement Learning for Single Humanoid Robot Locomotion Humanplus: Humanoid shadowing and imitation from humans,

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:30:30.640690Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-05T20:30:29.738800Z digest=sha256:a830b3395b42943c24acd53dff62ce722501f492e83a22fcf57ee579fa7c11c2

Observation 2e493578-05f6-4ac4-85d4-c5190a5de706 · outbound

This paper cites Okami: Teaching humanoid robots manipulation skills through single video imitation,.

MASH: Cooperative-Heterogeneous Multi-Agent Reinforcement Learning for Single Humanoid Robot Locomotion Okami: Teaching humanoid robots manipulation skills through single video imitation,

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:30:30.624099Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-05T20:30:29.880304Z digest=sha256:e1f95ca74f73aa1d0a32607fa348020ab66b5e229d6788668c80af232cdadadb

Observation 4883b107-7c53-4dc5-b71c-e1776702102c · outbound

This paper cites OmniH2O: Universal and dexterous human-to- humanoid whole-body teleoperation and learning,.

MASH: Cooperative-Heterogeneous Multi-Agent Reinforcement Learning for Single Humanoid Robot Locomotion OmniH2O: Universal and dexterous human-to- humanoid whole-body teleoperation and learning,

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:30:30.606108Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-05T20:30:29.953977Z digest=sha256:5ebc45c9f09439a6808cf4e7a6a3d4135a55e3734766761e651a20d45678ebe9

Observation bd3bae68-5775-4fe0-afac-cf3684981887 · outbound

This paper cites Perpetual humanoid control for real-time simulated avatars,.

MASH: Cooperative-Heterogeneous Multi-Agent Reinforcement Learning for Single Humanoid Robot Locomotion Perpetual humanoid control for real-time simulated avatars,

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-05T20:30:30.035721Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T20:30:30.035721Z digest=sha256:703ed3136eda454a8ab64aa7f0eed161efff3fdcac7333cb27a583e48f3ee035

Observation 12984db4-c843-4ea9-8300-f16881223180 · outbound

This paper cites Robust and versatile bipedal jumping control through reinforcement learning,.

MASH: Cooperative-Heterogeneous Multi-Agent Reinforcement Learning for Single Humanoid Robot Locomotion Robust and versatile bipedal jumping control through reinforcement learning,

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:30:30.576370Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-05T20:30:30.113842Z digest=sha256:2c3cb0c40fd56c2401e547caf0dd065bf14e550aaab17c98da3cf460ea6d4f6a

Observation d69fa8f1-b5db-4d51-9298-7c8e8479ace1 · outbound

This paper cites Feedback control for cassie with deep reinforcement learning,.

MASH: Cooperative-Heterogeneous Multi-Agent Reinforcement Learning for Single Humanoid Robot Locomotion Feedback control for cassie with deep reinforcement learning,

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:30:30.560029Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-05T20:30:30.119322Z digest=sha256:989015af56502ea67298c4f57ee22aee7dea2e8106264c072827116860085155

Observation 069b8b47-abbf-4977-ae2a-80882c267943 · outbound

This paper cites Sim-to-real learning of all common bipedal gaits via periodic reward composition,.

MASH: Cooperative-Heterogeneous Multi-Agent Reinforcement Learning for Single Humanoid Robot Locomotion Sim-to-real learning of all common bipedal gaits via periodic reward composition,

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:30:30.543052Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-05T20:30:30.125124Z digest=sha256:f863f69ff262fdfcf81edf667d7c240adff429e4150880e8bacfaecefa0cf350

Observation a0efdb97-5646-43d2-8fc7-8d0310bd2377 · outbound

This paper cites Amp: Adversarial motion priors for stylized physics-based character control,.

MASH: Cooperative-Heterogeneous Multi-Agent Reinforcement Learning for Single Humanoid Robot Locomotion Amp: Adversarial motion priors for stylized physics-based character control,

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:30:30.523071Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-05T20:30:30.131020Z digest=sha256:42dc4d45c9f55da8e7df9cc6c0018b657f87d0e4223620a7ea6d48e847533234

Observation d23ade03-2480-4779-9685-53f1e942aa9c · outbound

This paper cites Advancing humanoid locomotion: Mastering challenging terrains with denoising world model learning,.

MASH: Cooperative-Heterogeneous Multi-Agent Reinforcement Learning for Single Humanoid Robot Locomotion Advancing humanoid locomotion: Mastering challenging terrains with denoising world model learning,

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:30:30.501610Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-05T20:30:30.136272Z digest=sha256:8222ac7ad7f90817811417e6af8c4b9d1b9a6b24dee3b55f94799d7a5f8da17e

Observation b35ec145-653e-4fa0-8472-c6763c42a85b · outbound

This paper cites Reinforcement learning for swarm robotics: An overview of applications, algorithms and simulators,.

MASH: Cooperative-Heterogeneous Multi-Agent Reinforcement Learning for Single Humanoid Robot Locomotion Reinforcement learning for swarm robotics: An overview of applications, algorithms and simulators,

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:30:30.480618Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-05T20:30:30.142441Z digest=sha256:d6ce819ba47c1bd47fa4f855dff8616a91c33d9977da91970a5a5b6d443d11ec

Observation 5d7d7353-b146-4285-a9bf-f050193c4721 · outbound

This paper cites Smarts: An open-source scalable multi-agent rl training school for autonomous driving,.

MASH: Cooperative-Heterogeneous Multi-Agent Reinforcement Learning for Single Humanoid Robot Locomotion Smarts: An open-source scalable multi-agent rl training school for autonomous driving,

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:30:30.460116Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-05T20:30:30.149327Z digest=sha256:798e6abd5cb89d78b45badc9626f03678e741811fea773d1e4c42454d0a9825d

Observation 68817526-329f-484e-bd46-8fcd099f7438 · outbound

This paper cites Optimal tethered-uav deployment in a2g communication networks: Multi-agent q-learning approach,.

MASH: Cooperative-Heterogeneous Multi-Agent Reinforcement Learning for Single Humanoid Robot Locomotion Optimal tethered-uav deployment in a2g communication networks: Multi-agent q-learning approach,

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:30:30.441939Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-05T20:30:30.155078Z digest=sha256:ec809cc83e7b8dbd0b7fa5c970859b35fc7a08b29757016cac4b8df053d8526c

Observation 14e78a57-7704-4df5-badf-01df49855e08 · outbound

This paper cites Multi-Agent Target Assignment and Path Finding for Intelligent Warehouse: A Cooperative Multi-Agent Deep Reinforcement Learning Perspective.

MASH: Cooperative-Heterogeneous Multi-Agent Reinforcement Learning for Single Humanoid Robot Locomotion Multi-Agent Target Assignment and Path Finding for Intelligent Warehouse: A Cooperative Multi-Agent Deep Reinforcement Learning Perspective

Reference 41

Resolution
verified exact
local_arxiv, observed 2026-08-05T20:30:30.294226Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-05T20:30:30.161628Z digest=sha256:0275c6b8e2bd505c0167c36a8c44d89eb104f4b899670ce297f84e069d97c50c

Observation d1d5d281-6f41-495d-b014-22f98124cffd · outbound

This paper cites an unresolved cited work.

MASH: Cooperative-Heterogeneous Multi-Agent Reinforcement Learning for Single Humanoid Robot Locomotion Unresolved cited work

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-05T20:30:30.169226Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T20:30:30.169226Z digest=sha256:57f3e2cf26de961d7fcde3e0cfad0e9357db628ecbc80c727769697aa11599cf

Observation b0364da9-194a-4e12-940b-d733d99dcb95 · outbound

This paper cites Proximal Policy Optimization Algorithms.

MASH: Cooperative-Heterogeneous Multi-Agent Reinforcement Learning for Single Humanoid Robot Locomotion Proximal Policy Optimization Algorithms

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-05T20:30:30.174196Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T20:30:30.174196Z digest=sha256:5241c34bcf762d749d2f04af87b2997c6a8a2db10a83aec277bf06a3f29889b7

Observation cbe5f669-2f7d-4588-8e75-28bf798b0af0 · outbound

This paper cites Pomdps for robotic tasks with mixed observability.

MASH: Cooperative-Heterogeneous Multi-Agent Reinforcement Learning for Single Humanoid Robot Locomotion Pomdps for robotic tasks with mixed observability

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:30:30.405800Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-05T20:30:30.180731Z digest=sha256:0c8a24a61f4b2bde4d32295cdb4490473da13af0a94a59a8f31682fbc0d23eb3

Observation 29ca6b55-c894-4014-b506-112b19d6c1eb · outbound

This paper cites The surprising effectiveness of ppo in cooperative multi-agent games,.

MASH: Cooperative-Heterogeneous Multi-Agent Reinforcement Learning for Single Humanoid Robot Locomotion The surprising effectiveness of ppo in cooperative multi-agent games,

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:30:30.381407Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-05T20:30:30.187585Z digest=sha256:a58caa7cdb40c6f20d9e1db2a35df49ef66de8221f067f5264799e632f608784

Observation 9dd80522-a6bc-4869-8701-4b789a0946cb · outbound

This paper cites Stabilising experience replay for deep multi-agent rein- forcement learning,.

MASH: Cooperative-Heterogeneous Multi-Agent Reinforcement Learning for Single Humanoid Robot Locomotion Stabilising experience replay for deep multi-agent rein- forcement learning,

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:30:30.362079Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-05T20:30:30.193423Z digest=sha256:8ea7e847dd0609bbef1799d697d567ab79950156579d079a120f046c6742c76e

Observation 90a0c123-743d-4d53-b95a-8c19013752fc · outbound

This paper cites Isaac gym: High performance gpu based physics simulation for robot learning,.

MASH: Cooperative-Heterogeneous Multi-Agent Reinforcement Learning for Single Humanoid Robot Locomotion Isaac gym: High performance gpu based physics simulation for robot learning,

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:30:30.344647Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-05T20:30:30.203171Z digest=sha256:77107856dd66dc72b90b37d39b4f5c63c51515d283a6934d47d9d179ee1d4a70

Pith citing papers

No inbound Pith citation observations are available.