Pith. sign in

Paper Citation Record · LEDGER

MASH: Cooperative-Heterogeneous Multi-Agent Reinforcement Learning for Single Humanoid Robot Locomotion

As of 14 August 2026, this Paper Citation Record lists 47 of 47 outbound references and 0 inbound Pith citation observations for arXiv:2508.10423.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2508.10423 v1

Coverage vector

measured 47 of 47 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-05T20:30:30.203171Z

measured 47 of 47 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-13T06:32:02.005865+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

47 of 47 outbound references displayed

  • verified exact2
  • verified fuzzy37
  • unresolved8
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 77978675-00ca-4485-8090-d5e4154448f3 · outbound

This paper cites HOVER: Versatile neural whole-body controller for humanoid robots,.

MASH: Cooperative-Heterogeneous Multi-Agent Reinforcement Learning for Single Humanoid Robot Locomotion HOVER: Versatile neural whole-body controller for humanoid robots,

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:30:31.115929Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-05T20:30:26.580920Z digest=sha256:e34b78e8c888c8b4f024d97d6904e2cb2b58f25c9b387e2a75c36304fa18827b

Observation a7977c1f-a96e-4aa1-8d43-42653518fb8a · outbound

This paper cites Distributional policy gradient with distributional value function,.

MASH: Cooperative-Heterogeneous Multi-Agent Reinforcement Learning for Single Humanoid Robot Locomotion Distributional policy gradient with distributional value function,

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:30:31.098429Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-05T20:30:26.649591Z digest=sha256:5edc3b2d55e9827f967ed2ca3a062f098c866f45f66b0a04836e4e705d9f4fd7

Observation fe4023d4-eb8b-4e26-a495-6bf5f45352f6 · outbound

This paper cites Walk these ways: Tuning robot control for generalization with multiplicity of behavior,.

MASH: Cooperative-Heterogeneous Multi-Agent Reinforcement Learning for Single Humanoid Robot Locomotion Walk these ways: Tuning robot control for generalization with multiplicity of behavior,

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:30:31.081654Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-05T20:30:26.738893Z digest=sha256:c689acb18c6d14f17012f20b1bf4284e1f797b85bc0672668815982d58f9c341

Observation 985e80ca-683d-4f4e-bf17-4fa640ee2d8b · outbound

This paper cites Dreamwaq: Learning robust quadrupedal locomotion with implicit terrain imagination via deep reinforcement learning,.

MASH: Cooperative-Heterogeneous Multi-Agent Reinforcement Learning for Single Humanoid Robot Locomotion Dreamwaq: Learning robust quadrupedal locomotion with implicit terrain imagination via deep reinforcement learning,

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:30:31.063656Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-05T20:30:26.795999Z digest=sha256:85ea3d95df1cae401ea9ecffaf446045a1a8ce5a55a84ab04be6a55d644bd199

Observation 7c9e068d-c598-4426-a027-ebc4b4664807 · outbound

This paper cites Biped dynamic walking using reinforcement learning,.

MASH: Cooperative-Heterogeneous Multi-Agent Reinforcement Learning for Single Humanoid Robot Locomotion Biped dynamic walking using reinforcement learning,

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:30:31.045484Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-05T20:30:26.847516Z digest=sha256:b37ebf9ebc496350bf8be9e796493fb0ca8e0b2cbd9d479e84790211914f3e78

Observation c396c3dd-413b-41e3-8349-abb58f1d29e6 · outbound

This paper cites Learning vision-based bipedal locomotion for challeng- ing terrain,.

MASH: Cooperative-Heterogeneous Multi-Agent Reinforcement Learning for Single Humanoid Robot Locomotion Learning vision-based bipedal locomotion for challeng- ing terrain,

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:30:31.024478Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-05T20:30:26.945183Z digest=sha256:4ee9c564c5cc9a4d34ffa3a224fdbd6e3533cb05803122597b9ca4d441e04ed1

Observation 0ee0e265-40be-47bb-83e0-2ab770211a81 · outbound

This paper cites Motor anomaly detection for unmanned aerial vehicles using reinforcement learning,.

MASH: Cooperative-Heterogeneous Multi-Agent Reinforcement Learning for Single Humanoid Robot Locomotion Motor anomaly detection for unmanned aerial vehicles using reinforcement learning,

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:30:30.993595Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-05T20:30:27.019704Z digest=sha256:927dd0242f4a49cf1b88f7c258354c8a4cbf1177bffac9cfe3ef3121d92b9ffb

Observation c919e3cd-cf6a-4840-a824-a510f9f64f00 · outbound

This paper cites Optimization-based control for dynamic legged robots,.

MASH: Cooperative-Heterogeneous Multi-Agent Reinforcement Learning for Single Humanoid Robot Locomotion Optimization-based control for dynamic legged robots,

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:30:30.972502Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-05T20:30:27.111040Z digest=sha256:60aeb2d899b4d40c02bf1e6e077cb8d40a2c1e3c598174b793be6a1e1c9aa4fd

Observation 6b8305bd-9467-4bbb-a3b7-aca9f7e45c64 · outbound

This paper cites Versatile multicontact planning and control for legged loco-manipulation,.

MASH: Cooperative-Heterogeneous Multi-Agent Reinforcement Learning for Single Humanoid Robot Locomotion Versatile multicontact planning and control for legged loco-manipulation,

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-05T20:30:27.279483Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T20:30:27.279483Z digest=sha256:c94526575e3026c50b51efa479fc399451d8dbb34e6b77d2acf068da36e6dbb6

Observation 34a470d6-66e6-426a-89b5-32f07bea1041 · outbound

This paper cites Combining trajectory optimization, supervised machine learning, and model structure for mitigating the curse of dimensionality in the control of bipedal robots,.

MASH: Cooperative-Heterogeneous Multi-Agent Reinforcement Learning for Single Humanoid Robot Locomotion Combining trajectory optimization, supervised machine learning, and model structure for mitigating the curse of dimensionality in the control of bipedal robots,

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-05T20:30:27.398322Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T20:30:27.398322Z digest=sha256:3c1ac70f13483f5199d3e2ad6bcd23fee36b1fecbf6270c947ab8681e5cf413c

Observation db20e7da-e494-42b6-9fa5-c72d95d9758c · outbound

This paper cites Real-world humanoid locomotion with reinforcement learning,.

MASH: Cooperative-Heterogeneous Multi-Agent Reinforcement Learning for Single Humanoid Robot Locomotion Real-world humanoid locomotion with reinforcement learning,

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-05T20:30:27.567294Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T20:30:27.567294Z digest=sha256:b56443ab4781cab7558857c665d82c7c6a33d47192ac591051858abbc2e72abd

Observation 796fbf8d-98bb-459e-9f76-c6f243150cdc · outbound

This paper cites Not only rewards but also constraints: Applications on legged robot locomotion,.

MASH: Cooperative-Heterogeneous Multi-Agent Reinforcement Learning for Single Humanoid Robot Locomotion Not only rewards but also constraints: Applications on legged robot locomotion,

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-05T20:30:27.661298Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T20:30:27.661298Z digest=sha256:63c3b4fe1a7f1d6bb150951c48ca8f1ede56bd567655bd3400da8d6edce39881

Observation 4c7eb530-016d-42ac-9dea-1a50c79ae653 · outbound

This paper cites Diffusion policy: Visuomotor policy learning via ac- tion diffusion,.

MASH: Cooperative-Heterogeneous Multi-Agent Reinforcement Learning for Single Humanoid Robot Locomotion Diffusion policy: Visuomotor policy learning via ac- tion diffusion,

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-05T20:30:27.776297Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T20:30:27.776297Z digest=sha256:25c02f0819ec7546df1cf977360aa17ef6e223b53161da04e824ea9112ebe8c3

Observation deca8a10-e073-4628-953f-b38508e1f524 · outbound

This paper cites Learning-based legged locomotion: State of the art and future perspectives,.

MASH: Cooperative-Heterogeneous Multi-Agent Reinforcement Learning for Single Humanoid Robot Locomotion Learning-based legged locomotion: State of the art and future perspectives,

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:30:30.888448Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-05T20:30:27.906470Z digest=sha256:874c65cc91066128a8f546fcad022a086eab4c41924bc8dbdfd1e3d0f3ea3824

Observation 4becff69-4421-4263-b090-a7df74190e7a · outbound

This paper cites Learning whole-body loco-manipulation for omni-directional task space pose tracking with a wheeled-quadrupedal-manipulator,.

MASH: Cooperative-Heterogeneous Multi-Agent Reinforcement Learning for Single Humanoid Robot Locomotion Learning whole-body loco-manipulation for omni-directional task space pose tracking with a wheeled-quadrupedal-manipulator,

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:30:30.870338Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-05T20:30:28.024263Z digest=sha256:80dc2a812dfda3d929fca07fae58d849185a91014120b34dbfae4738985d080d

Observation 46252b04-f802-4a0a-88e7-4d329b2778cf · outbound

This paper cites Learning agile soccer skills for a bipedal robot with deep reinforcement learning,.

MASH: Cooperative-Heterogeneous Multi-Agent Reinforcement Learning for Single Humanoid Robot Locomotion Learning agile soccer skills for a bipedal robot with deep reinforcement learning,

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:30:30.853551Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-05T20:30:28.188458Z digest=sha256:fbc596b3ea30776b5dd349419b255f97051d4ec55c2f105b7624c755e8e3519a

Observation cd324124-89f6-4c88-95c0-b2eacca536ee · outbound

This paper cites Visual whole-body control for legged loco-manipulation,.

MASH: Cooperative-Heterogeneous Multi-Agent Reinforcement Learning for Single Humanoid Robot Locomotion Visual whole-body control for legged loco-manipulation,

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:30:30.837324Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-05T20:30:28.307712Z digest=sha256:f4ba2bd63e8664340dade52638bc1399804259e927a158c0a63887e6ddaa0d30

Observation bbbe3bc7-956b-48c5-b565-3979b034076e · outbound

This paper cites Teleoperation of humanoid robots: A survey,.

MASH: Cooperative-Heterogeneous Multi-Agent Reinforcement Learning for Single Humanoid Robot Locomotion Teleoperation of humanoid robots: A survey,

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:30:30.822545Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-05T20:30:28.432305Z digest=sha256:d177ca3d9cadc5702af4ff0c8788e299f1df401df287dbb9089443a7d08598c2

Observation b80c870b-c2c3-448b-bcea-1eaf90a43018 · outbound

This paper cites Sim-to-real robotic sketching using behavior cloning and reinforcement learning,.

MASH: Cooperative-Heterogeneous Multi-Agent Reinforcement Learning for Single Humanoid Robot Locomotion Sim-to-real robotic sketching using behavior cloning and reinforcement learning,

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:30:30.805259Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-05T20:30:28.565492Z digest=sha256:2bdb0460fdeaf11d1482efa8b10ad9eac9c880aa1fa2da047f369bfc5b386a13

Observation a823bc1a-a4d1-4025-b640-27f1e7692334 · outbound

This paper cites A composite control strategy for quadruped robot by integrating reinforcement learning and model-based control,.

MASH: Cooperative-Heterogeneous Multi-Agent Reinforcement Learning for Single Humanoid Robot Locomotion A composite control strategy for quadruped robot by integrating reinforcement learning and model-based control,

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:30:30.787160Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-05T20:30:28.728125Z digest=sha256:f5370bd31a8da1f9b5ce3a5d47335aed1499a39a040d1df68cc2b4d5448f6e00

Observation edfb33f5-fe21-4ca7-b58f-03b78e567cab · outbound

This paper cites Towards human-level bimanual dexterous manipulation with reinforcement learning,.

MASH: Cooperative-Heterogeneous Multi-Agent Reinforcement Learning for Single Humanoid Robot Locomotion Towards human-level bimanual dexterous manipulation with reinforcement learning,

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:30:30.770371Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-05T20:30:28.852754Z digest=sha256:4070f3d4ddf7ee5c23223cb1d55deb39d3820a7fc54192e0e7e7442908fd382b

Observation 3ea2df64-31e5-4e72-b1f1-1387ac8df8af · outbound

This paper cites Monotonic value function factorisation for deep multi- agent reinforcement learning,.

MASH: Cooperative-Heterogeneous Multi-Agent Reinforcement Learning for Single Humanoid Robot Locomotion Monotonic value function factorisation for deep multi- agent reinforcement learning,

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:30:30.753663Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-05T20:30:29.025150Z digest=sha256:0c0df8efdefece033fa0f0853a234547f336ffd88e580859c8ef99582c98817e

Observation 7c4e3bfc-c2bd-4179-8d5b-6abc35ee01bf · outbound

This paper cites Data efficient deep reinforcement learning with action-ranked temporal difference learning,.

MASH: Cooperative-Heterogeneous Multi-Agent Reinforcement Learning for Single Humanoid Robot Locomotion Data efficient deep reinforcement learning with action-ranked temporal difference learning,

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:30:30.736318Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-05T20:30:29.167712Z digest=sha256:3257e31e7621cc378dd60203469fb3b24a52f00e2ac4ceb8d7b753d766a4de48

Observation 578cc9d9-97ad-47af-97ce-8f23987a7746 · outbound

This paper cites MASQ: Multi-Agent Reinforcement Learning for Single Quadruped Robot Locomotion.

MASH: Cooperative-Heterogeneous Multi-Agent Reinforcement Learning for Single Humanoid Robot Locomotion MASQ: Multi-Agent Reinforcement Learning for Single Quadruped Robot Locomotion

Reference 24

Resolution
verified exact
local_arxiv, observed 2026-08-05T20:30:30.324419Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-05T20:30:29.314929Z digest=sha256:5db9148bd5cfcdbb2e5651ee37d5189e9a5010a0d96da5439506874090790cbe

Observation e95550f3-e507-4b68-b817-530007f7ea10 · outbound

This paper cites Expressive whole-body control for humanoid robots,.

MASH: Cooperative-Heterogeneous Multi-Agent Reinforcement Learning for Single Humanoid Robot Locomotion Expressive whole-body control for humanoid robots,

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:30:30.719256Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-05T20:30:29.389280Z digest=sha256:bf66eee791efcf4645b7dbc19cddb4cb657a313d598490cc7b5e4873bc5f8a51

Observation 33c3c267-bcbf-482c-ac3e-13ba11571aa4 · outbound

This paper cites Mobile-television: Predictive motion priors for hu- manoid whole-body control,.

MASH: Cooperative-Heterogeneous Multi-Agent Reinforcement Learning for Single Humanoid Robot Locomotion Mobile-television: Predictive motion priors for hu- manoid whole-body control,

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:30:30.697923Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-05T20:30:29.460025Z digest=sha256:bbd8bb692bbbf01f606a5a32ec90432c162f1175c9d4a035f376638dcef5c191

Observation c211d33c-216c-4a26-977a-e24b062e3fac · outbound

This paper cites Wococo: Learning whole-body humanoid control with sequential contacts,.

MASH: Cooperative-Heterogeneous Multi-Agent Reinforcement Learning for Single Humanoid Robot Locomotion Wococo: Learning whole-body humanoid control with sequential contacts,

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:30:30.677407Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-05T20:30:29.550583Z digest=sha256:791ecb45e6f8d2f644aa98300fcf1dc396aace7b90c931e03e1399043bf8b1db

Observation 58e79e0c-8b28-4c1c-a432-294f41703d36 · outbound

This paper cites Learning human-to-humanoid real-time whole-body teleoperation,.

MASH: Cooperative-Heterogeneous Multi-Agent Reinforcement Learning for Single Humanoid Robot Locomotion Learning human-to-humanoid real-time whole-body teleoperation,

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:30:30.659984Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-05T20:30:29.637244Z digest=sha256:0e2260e7413206c5d53c253747205cac450a6c1a54021ad7e30427b339baced2

Observation a15a4892-8a69-4a42-8efe-d4763eda588c · outbound

This paper cites Humanplus: Humanoid shadowing and imitation from humans,.

MASH: Cooperative-Heterogeneous Multi-Agent Reinforcement Learning for Single Humanoid Robot Locomotion Humanplus: Humanoid shadowing and imitation from humans,

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:30:30.640690Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-05T20:30:29.738800Z digest=sha256:eedbfc5a7340248c2267cf402c3b402da5ed2179ccf0eaaf04d2d60f2fac019c

Observation 2e493578-05f6-4ac4-85d4-c5190a5de706 · outbound

This paper cites Okami: Teaching humanoid robots manipulation skills through single video imitation,.

MASH: Cooperative-Heterogeneous Multi-Agent Reinforcement Learning for Single Humanoid Robot Locomotion Okami: Teaching humanoid robots manipulation skills through single video imitation,

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:30:30.624099Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-05T20:30:29.880304Z digest=sha256:f9f71afcf321d714031341f7b364867455c2a25b3af418ebb49c2ab9fbd253dc

Observation 4883b107-7c53-4dc5-b71c-e1776702102c · outbound

This paper cites OmniH2O: Universal and dexterous human-to- humanoid whole-body teleoperation and learning,.

MASH: Cooperative-Heterogeneous Multi-Agent Reinforcement Learning for Single Humanoid Robot Locomotion OmniH2O: Universal and dexterous human-to- humanoid whole-body teleoperation and learning,

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:30:30.606108Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-05T20:30:29.953977Z digest=sha256:111bc7ea0d32d0344f56e4175094e6fb17d6c2bdaaa24006ccdd286823f399f2

Observation bd3bae68-5775-4fe0-afac-cf3684981887 · outbound

This paper cites Perpetual humanoid control for real-time simulated avatars,.

MASH: Cooperative-Heterogeneous Multi-Agent Reinforcement Learning for Single Humanoid Robot Locomotion Perpetual humanoid control for real-time simulated avatars,

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-05T20:30:30.035721Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T20:30:30.035721Z digest=sha256:703ed3136eda454a8ab64aa7f0eed161efff3fdcac7333cb27a583e48f3ee035

Observation 12984db4-c843-4ea9-8300-f16881223180 · outbound

This paper cites Robust and versatile bipedal jumping control through reinforcement learning,.

MASH: Cooperative-Heterogeneous Multi-Agent Reinforcement Learning for Single Humanoid Robot Locomotion Robust and versatile bipedal jumping control through reinforcement learning,

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:30:30.576370Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-05T20:30:30.113842Z digest=sha256:2966196c7fa58025659b29c75dc676ab97a08f43dc3c844d41a4b2def3662973

Observation d69fa8f1-b5db-4d51-9298-7c8e8479ace1 · outbound

This paper cites Feedback control for cassie with deep reinforcement learning,.

MASH: Cooperative-Heterogeneous Multi-Agent Reinforcement Learning for Single Humanoid Robot Locomotion Feedback control for cassie with deep reinforcement learning,

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:30:30.560029Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-05T20:30:30.119322Z digest=sha256:036813f423f3e63f829dbc9171ad7f890c0fc2e633fae708335312da28c71ab2

Observation 069b8b47-abbf-4977-ae2a-80882c267943 · outbound

This paper cites Sim-to-real learning of all common bipedal gaits via periodic reward composition,.

MASH: Cooperative-Heterogeneous Multi-Agent Reinforcement Learning for Single Humanoid Robot Locomotion Sim-to-real learning of all common bipedal gaits via periodic reward composition,

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:30:30.543052Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-05T20:30:30.125124Z digest=sha256:aa04d9b1341ba09f8f2e57191f84801531f3264f62775faeae9363684a7afcca

Observation a0efdb97-5646-43d2-8fc7-8d0310bd2377 · outbound

This paper cites Amp: Adversarial motion priors for stylized physics-based character control,.

MASH: Cooperative-Heterogeneous Multi-Agent Reinforcement Learning for Single Humanoid Robot Locomotion Amp: Adversarial motion priors for stylized physics-based character control,

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:30:30.523071Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-05T20:30:30.131020Z digest=sha256:3c26e5d01278c94a50ac93d541e511f08932aec9a295acdb22f7dcde6b0380a5

Observation d23ade03-2480-4779-9685-53f1e942aa9c · outbound

This paper cites Advancing humanoid locomotion: Mastering challenging terrains with denoising world model learning,.

MASH: Cooperative-Heterogeneous Multi-Agent Reinforcement Learning for Single Humanoid Robot Locomotion Advancing humanoid locomotion: Mastering challenging terrains with denoising world model learning,

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:30:30.501610Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-05T20:30:30.136272Z digest=sha256:297d3d09ed915507236d0a3fb5dcec015760e3a725bc9a62667c46100861a119

Observation b35ec145-653e-4fa0-8472-c6763c42a85b · outbound

This paper cites Reinforcement learning for swarm robotics: An overview of applications, algorithms and simulators,.

MASH: Cooperative-Heterogeneous Multi-Agent Reinforcement Learning for Single Humanoid Robot Locomotion Reinforcement learning for swarm robotics: An overview of applications, algorithms and simulators,

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:30:30.480618Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-05T20:30:30.142441Z digest=sha256:9f612949d6ead98557887e5a0fb8667e10019f141f51de89b5d733d6b114a493

Observation 5d7d7353-b146-4285-a9bf-f050193c4721 · outbound

This paper cites Smarts: An open-source scalable multi-agent rl training school for autonomous driving,.

MASH: Cooperative-Heterogeneous Multi-Agent Reinforcement Learning for Single Humanoid Robot Locomotion Smarts: An open-source scalable multi-agent rl training school for autonomous driving,

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:30:30.460116Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-05T20:30:30.149327Z digest=sha256:42641099878f5b981ea139bd3779300de1d7a491768a5221a60818dd12343c62

Observation 68817526-329f-484e-bd46-8fcd099f7438 · outbound

This paper cites Optimal tethered-uav deployment in a2g communication networks: Multi-agent q-learning approach,.

MASH: Cooperative-Heterogeneous Multi-Agent Reinforcement Learning for Single Humanoid Robot Locomotion Optimal tethered-uav deployment in a2g communication networks: Multi-agent q-learning approach,

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:30:30.441939Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-05T20:30:30.155078Z digest=sha256:37fb66c9011149f4f8c236fedc6b486f53082083d6de34771000dd4605368443

Observation 14e78a57-7704-4df5-badf-01df49855e08 · outbound

This paper cites Multi-Agent Target Assignment and Path Finding for Intelligent Warehouse: A Cooperative Multi-Agent Deep Reinforcement Learning Perspective.

MASH: Cooperative-Heterogeneous Multi-Agent Reinforcement Learning for Single Humanoid Robot Locomotion Multi-Agent Target Assignment and Path Finding for Intelligent Warehouse: A Cooperative Multi-Agent Deep Reinforcement Learning Perspective

Reference 41

Resolution
verified exact
local_arxiv, observed 2026-08-05T20:30:30.294226Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-05T20:30:30.161628Z digest=sha256:449ea7ce82a06a875b7458fc954c963ea8fb9acbdab7d2f00a92403e96879359

Observation d1d5d281-6f41-495d-b014-22f98124cffd · outbound

This paper cites an unresolved cited work.

MASH: Cooperative-Heterogeneous Multi-Agent Reinforcement Learning for Single Humanoid Robot Locomotion Unresolved cited work

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-05T20:30:30.169226Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T20:30:30.169226Z digest=sha256:57f3e2cf26de961d7fcde3e0cfad0e9357db628ecbc80c727769697aa11599cf

Observation b0364da9-194a-4e12-940b-d733d99dcb95 · outbound

This paper cites Proximal Policy Optimization Algorithms.

MASH: Cooperative-Heterogeneous Multi-Agent Reinforcement Learning for Single Humanoid Robot Locomotion Proximal Policy Optimization Algorithms

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-05T20:30:30.174196Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T20:30:30.174196Z digest=sha256:a982c6fb3e5239a38512c5fc5ecaabf7d23a42d8980bcd587308edb5f78193e5

Observation cbe5f669-2f7d-4588-8e75-28bf798b0af0 · outbound

This paper cites Pomdps for robotic tasks with mixed observability.

MASH: Cooperative-Heterogeneous Multi-Agent Reinforcement Learning for Single Humanoid Robot Locomotion Pomdps for robotic tasks with mixed observability

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:30:30.405800Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-05T20:30:30.180731Z digest=sha256:75d7fcdfa62c737bf321b496bb983516cb5565bfa81c1df1e8297f36c874a7cc

Observation 29ca6b55-c894-4014-b506-112b19d6c1eb · outbound

This paper cites The surprising effectiveness of ppo in cooperative multi-agent games,.

MASH: Cooperative-Heterogeneous Multi-Agent Reinforcement Learning for Single Humanoid Robot Locomotion The surprising effectiveness of ppo in cooperative multi-agent games,

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:30:30.381407Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-05T20:30:30.187585Z digest=sha256:73d7ea49a4ce3c78b24232381715e0518e285f6c865b51cc7d446ae27616dd5f

Observation 9dd80522-a6bc-4869-8701-4b789a0946cb · outbound

This paper cites Stabilising experience replay for deep multi-agent rein- forcement learning,.

MASH: Cooperative-Heterogeneous Multi-Agent Reinforcement Learning for Single Humanoid Robot Locomotion Stabilising experience replay for deep multi-agent rein- forcement learning,

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:30:30.362079Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-05T20:30:30.193423Z digest=sha256:fe9094f410b65a2b84accd7e2ef9fa9160f392b6829b62eb81dbb72a98724725

Observation 90a0c123-743d-4d53-b95a-8c19013752fc · outbound

This paper cites Isaac gym: High performance gpu based physics simulation for robot learning,.

MASH: Cooperative-Heterogeneous Multi-Agent Reinforcement Learning for Single Humanoid Robot Locomotion Isaac gym: High performance gpu based physics simulation for robot learning,

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:30:30.344647Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-05T20:30:30.203171Z digest=sha256:43a5c72b79d59b6c0160437c7a57e19ffbdb314d14821964858fe28075680399

Pith citing papers

No inbound Pith citation observations are available.