REVIEW 3 major objections 2 minor 38 references
Bridging the Sim-to-Real Gap in Parallel-Link Leg Mechanisms via Simulator-Side Dynamics Normalization
T0 review · 3 major / 2 minor · reviewed 2026-08-07 · deepseek-v4-flash
Pith's one-line read This paper claims that the sim-to-real gap for parallel-link legs comes from missing actuator and linkage dynamics in the serial-tree surrogate, and that adding them on the simulator side cuts joint-position and torque errors by roughly…
desk verdict Plausible and useful method for a real gap, but the headline numbers depend on identification details the abstract doesn't show; worth a rigorous review. read the letter →
The pith
A machine-rendered reading of the paper's core claim, the machinery that carries it, and where it could break.
The reading
What carries the argument
The load-bearing object is the coordinate transformation that maps physical parallel-link actuator and leg dynamics into the serial-tree simulator's joint space. S3N-Act uses this transformation to add actuator inertia and damping; S3N-Full goes further, using separately identified actuator-level and leg-level frequency responses to restore residual linkage inertia that the serial-tree reduction drops. The construction preserves the tree topology, which is what allows a standard serial-link simulator and policy trainer to run unchanged while the effective dynamics match the physical parallel-link leg.
What would settle it
Identify the actuator- and leg-level frequency responses on the real leg, then command motions at frequencies and torque amplitudes outside the identification range; if the normalized simulator's predicted joint positions, torques, or ground reaction forces deviate from hardware measurements by much more than the reported 9.9% gap, the linear additive normalization is not transferable. A simpler check is to repeat the identification at two excitation amplitudes and see whether the estimated inertia and damping change.
Extended reading notes
Core claim
The central claim is that the sim-to-real error caused by replacing a parallel-link mechanism with a serial-tree surrogate is mostly a dynamics-normalization problem, not just a kinematic-mapping problem. Conventional Jacobian-based mappings preserve kinematic consistency and virtual-work relations but leave actuator inertia, actuator damping, and omitted linkage inertia in the wrong coordinates. S3N restores these by transforming actuator inertia and damping into serial coordinates (S3N-Act) and, in S3N-Full, adding leg-level residuals identified from actuator- and leg-level frequency responses. Measured against a Jacobian-mapping baseline, the full method reduces joint-position RMSE by 80.9% and torque RMSE by 82.1%, lowers ground-reaction-force RMSE by 62% to 65% in pitch-in-place motion, and brings the command-normalized sim-to-real gap down from 17.3% to 9.9% in circular locomotion.
Load-bearing premise
The load-bearing premise is that the real leg's actuator and linkage behavior can be represented as additive inertia and damping measured through frequency responses; if the hardware has significant friction, backlash, or structural flexibility that those linear measurements do not capture, the normalized simulator will not match the real robot outside the identified operating range.
Editorial extensions
If this is right
- A serial-tree simulator can be made hardware-consistent for a parallel-link leg by calibrating dynamics on the simulator side, without switching to a closed-loop linkage simulation.
- Both motion-level metrics (joint position, torque) and force-level metrics (ground reaction force) improve under S3N, so policies trained in the normalized simulator should transfer better in contact-rich tasks.
- With the full identification, the sim-to-real gap in circular locomotion drops from 17.3% to 9.9%, implying that usable low-error training can happen in the serial-tree framework.
- Because the method relies on measured frequency responses rather than mechanism-specific analytic dynamics, the same normalization recipe can be applied to other parallel-link mechanisms.
Reading between the lines
- A natural extension the paper does not spell out is that the identified normalization parameters could be re-estimated quickly when hardware changes (new actuators, added payload, wear), making the calibration a maintenance step rather than a one-time fix.
- The frequency-response basis implies a testable boundary: the normalization should stay valid only inside the frequency and amplitude range of the identification excitation, so excitation design is likely the practical limit of the method.
- The method could be combined with domain randomization by normalizing the nominal dynamics first and then randomizing around the normalized parameters, which would preserve the measured hardware behavior while retaining robustness to unmodeled variation.
Editorial analysis
A structured set of objections, weighed in public.
Referee Report
Summary. The paper proposes a simulator-side dynamics normalization approach (S3N) for parallel-link leg mechanisms simulated via a serial-tree surrogate. It introduces two variants: S3N-Act, which maps actuator inertia and damping into serial coordinates via coordinate transformation, and S3N-Full, which additionally identifies actuator- and leg-level frequency responses to restore residual linkage inertia. On a 2-DoF validation, S3N-Full reduces joint-position and torque RMSEs by 80.9% and 82.1% relative to a Jacobian-mapping baseline, and reduces the phase-averaged, command-normalized sim-to-real gap during circular locomotion from 17.3% to 9.9%.
Significance. If the reported results are reproducible, the contribution is practically significant: it offers a way to train policies in a serial-tree simulator while preserving the dynamics of the physical parallel-link mechanism, without requiring a full parallel-link simulation. The paper's explicit targeting of both motion- and force-level consistency is a strength, as is its use of multiple metrics (joint position, torque, ground reaction force). However, the abstract alone does not provide enough detail to verify the identification procedure or the statistical basis of the improvement claims; the central claims are plausible but currently unverified.
major comments (3)
- [Abstract / S3N-Full identification] The central result of the paper rests on the S3N-Full identification step, which is stated only as 'separately identifying actuator- and leg-level frequency responses.' The abstract does not provide the identification excitation, the frequency range, the model structure (e.g., order of the frequency-response fit), or any demonstration that the identified linear models capture the nonlinearities (friction, backlash, joint flexibility) of the physical leg. Because the improvement claims (80.9% joint-position and 82.1% torque RMSE reduction) are the paper's main quantitative evidence, the absence of this derivation and its validation is a load-bearing gap that must be addressed.
- [Abstract / experimental methodology] All reported improvements are point estimates without error bars, trial counts, or statistical tests. In particular, the phase-averaged, command-normalized sim-to-real gap reduction from 17.3% to 9.9% is a single number with no variance or number of trials; it is impossible to distinguish a systematic effect from run-to-run variability. The authors must provide the number of independent runs, the standard deviation or confidence interval, and the exact definition of the phase-averaged and command-normalized metrics.
- [Abstract / generalization across workspace] For a parallel-link mechanism, the inertia reflected to each actuator depends on the configuration through the mechanism Jacobian. The paper claims that S3N maps identified inertia/damping parameters into serial-tree coordinates as additive terms, but the abstract does not establish that these identified parameters are invariant across the workspace or that the test trajectories (pitch-in-place, circular locomotion) lie within the identification excitation envelope. If the identification was performed at a single pose or with small-amplitude excitation, the reported RMSE reductions could reflect overfitting to a narrow operating region rather than a general dynamics correction. The authors should report the workspace coverage of the identification data and validate the normalization on motions that are outside the identification set.
minor comments (2)
- [Abstract / terminology] The term 'phase-averaged, command-normalized sim-to-real gap' is insufficiently defined in the abstract; a formula or a reference should be provided so the metric is reproducible.
- [Abstract / platform details] The abstract does not state the robot platform details (dimensions, actuators, control frequency) or the simulator used; these details are needed for reproducibility and should be mentioned or cited.
Circularity Check
No specific circular step is exhibited; S3N's fitted parameters are validated on separate motion tasks, and no equation in the abstract reduces a prediction to its fit.
full rationale
The abstract's derivation chain consists of three distinguishable parts: (i) an analytic coordinate transformation that maps actuator inertia and damping into the serial-tree simulator coordinates (S3N-Act), (ii) an empirical identification of actuator- and leg-level frequency responses that supplies residual linkage inertia (S3N-Full), and (iii) an evaluation of the resulting simulator on pitch-in-place and circular-locomotion tasks. Parts (i) and (ii) are model-construction steps; they do not use the reported RMSE or gap-reduction numbers as inputs. Part (iii) compares the normalized simulator against real-robot measurements on motions that are not stated to be the same signals used for identification. The abstract does not exhibit any equation in which a fitted parameter is algebraically identical to a reported prediction, nor does it invoke a self-citation as the justification for the central claim. A concern that the identification and evaluation datasets might overlap would be a legitimate reproducibility question, but under the hard rules circularity cannot be inferred from the absence of that information. Therefore no demonstrated circular step is present in the available text.
Assumptions & free parameters
free parameters (2)
- actuator inertia/damping parameters =
not reported
- residual linkage inertia parameters =
not reported
assumptions (3)
- domain assumption Serial-tree surrogate plus added inertia/damping can represent parallel-link dynamics
- domain assumption Actuator and leg dynamics are linear time-invariant over the operating range
- standard math Jacobian-based mappings correctly capture kinematic and virtual-work consistency
Cite this review
Pith. "Pith review of Bridging the Sim-to-Real Gap in Parallel-Link Leg Mechanisms via Simulator-Side Dynamics Normalization." pith.science (2026). https://pith.science/paper/VIJCVALE
@misc{pith2026260801697,
author = {Pith},
title = {Pith review of: Bridging the Sim-to-Real Gap in Parallel-Link Leg Mechanisms via Simulator-Side Dynamics Normalization},
year = {2026},
howpublished = {\url{https://pith.science/paper/VIJCVALE}},
note = {Machine review of arXiv:2608.01697}
}
read the original abstract
This paper addresses the sim-to-real gap in dynamics arising when a parallel-link mechanism is represented by a serial-tree surrogate in simulation. Conventional Jacobian-based state and torque mappings preserve consistency with the kinematic and virtual-work relations but do not account for the coordinate-induced redistribution of actuator inertia and damping and the linkage inertia omitted during serial-tree reduction. To address this gap, Simulator-Side System Normalization (S3N) is proposed to normalize the serial-tree simulator's effective dynamics while preserving its tree topology. S3N-Act incorporates actuator inertia and damping into the serial-coordinate dynamics through coordinate transformation, whereas S3N-Full restores residual linkage inertia by separately identifying actuator- and leg-level frequency responses. In the 2-DoF validation, S3N-Full reduced the joint-position and torque RMSEs by 80.9% and 82.1%, respectively, relative to the Jacobian-mapping baseline. During pitch-in-place motion, S3N-Act and S3N-Full reduced the RMSE of the ground reaction force norm by 65.1% and 62.4%, respectively. During circular locomotion, S3N-Full reduced the phase-averaged, command-normalized sim-to-real gap from 17.3% to 9.9%. These results show that simulator-side normalization improves motion- and force-level sim-to-real consistency. It enables policy training in a serial-tree framework with hardware-consistent dynamics that better represent the physical parallel-link mechanism.
Figures
Figures from the paper (8 more)
Reference graph
Works this paper leans on
-
[1]
Merlet,Parallel Robots, 2nd ed
J.-P. Merlet,Parallel Robots, 2nd ed. Springer, 2006
2006
-
[2]
Briot and W
S. Briot and W. Khalil,Dynamics of Parallel Robots: From Rigid Bodies to Flexible Elements. Springer, 2015
2015
-
[3]
Mathematical and experi- mental verification of efficient force transmission by biarticular muscle actuator,
S. Oh, V . Salvucci, Y . Kimura, and Y . Hori, “Mathematical and experi- mental verification of efficient force transmission by biarticular muscle actuator,” inProceedings of the 18th IFAC World Congress, ser. IFAC Proceedings V olumes, vol. 44, no. 1. IFAC Secretariat, 2011, pp. 13 516–13 521
2011
-
[4]
A parallel actuated pantograph leg for high-speed locomotion,
W. Guo, C. Cai, M. Li, F. Zha, P. Wang, and K. Wang, “A parallel actuated pantograph leg for high-speed locomotion,”Journal of Bionic Engineering, vol. 14, no. 2, pp. 202–217, 2017. 10
work page 2017
-
[5]
Proprioceptive actuator design in the MIT Cheetah: Impact mitigation and high-bandwidth physical interaction for dynamic legged robots,
P. M. Wensing, A. Wang, S. Seok, D. Otten, J. Lang, and S. Kim, “Proprioceptive actuator design in the MIT Cheetah: Impact mitigation and high-bandwidth physical interaction for dynamic legged robots,” IEEE Transactions on Robotics, vol. 33, no. 3, pp. 509–522, 2017
2017
-
[6]
E. M. Hoffman, A. Curti, N. Miguel, S. K. Kothakota, A. Molina, A. Roig, and L. Marchionni, “Modeling and numerical analysis of Kangaroo lower body based on constrained dynamics of hybrid serial– parallel floating-base systems,”Robotics and Autonomous Systems, vol. 182, p. 104827, 2024, arXiv:2312.04161
arXiv 2024
-
[7]
SLIP embodied robust quadruped robot control,
J. Hong, C. Yeo, S. Bae, J. Hong, and S. Oh, “SLIP embodied robust quadruped robot control,” in2024 IEEE/RSJ International Conference on Intelligent Robots and Systems (IROS), 2024, pp. 14 219–14 224
work page 2024
-
[8]
Y . Tanaka, A. Zhu, Q. Wang, and D. Hong, “Mechanical intelligence- aware curriculum reinforcement learning for humanoids with parallel actuation,” in2025 IEEE-RAS 24th International Conference on Hu- manoid Robots (Humanoids), 2025, pp. 882–889
work page 2025
Show all 38 references
-
[9]
High dynamic position control for a typical hydraulic quadruped robot leg based on virtual decomposition control,
K. Zhang, J. Zhang, H. Zong, L. Fang, J. Shen, M. Cheng, and B. Xu, “High dynamic position control for a typical hydraulic quadruped robot leg based on virtual decomposition control,”IEEE/ASME Transactions on Mechatronics, vol. 30, no. 4, pp. 2473–2484, Aug. 2025
2025
-
[10]
Impedance matching: Enabling an RL-based running jump in a quadruped robot,
N. Guan, S. Yu, S. Zhu, and D. Kim, “Impedance matching: Enabling an RL-based running jump in a quadruped robot,” in2024 21st Inter- national Conference on Ubiquitous Robots (UR), 2024, pp. 755–761
2024
-
[11]
MuJoCo: A physics engine for model-based control,
E. Todorov, T. Erez, and Y . Tassa, “MuJoCo: A physics engine for model-based control,” inProceedings of the IEEE/RSJ International Conference on Intelligent Robots and Systems, 2012, pp. 5026–5033
2012
-
[12]
Learning to walk with hybrid serial– parallel linkages: A case study on the Kangaroo robot,
F. Amadio, H. Li, L. Uttini, D. Kanoulas, S. Ivaldi, V . Modugno, and E. M. Hoffman, “Learning to walk with hybrid serial– parallel linkages: A case study on the Kangaroo robot,” inResults from the 13th International Conference on Robot Intelligence Technology and Applications...
2025
-
[13]
MuJoCo documentation: Computation,
Google DeepMind, “MuJoCo documentation: Computation,” https://mujoco.readthedocs.io/en/stable/computation/index.html, accessed: 2026-03-19
2026
-
[14]
MuJoCo documentation: Modeling,
——, “MuJoCo documentation: Modeling,” https://mujoco.readthedocs. io/en/stable/modeling.html, accessed: 2026-03-19
2026
-
[15]
NimbRo-OP2: Grown- up 3D printed open humanoid platform for research,
G. Ficht, P. Allgeuer, H. Farazi, and S. Behnke, “NimbRo-OP2: Grown- up 3D printed open humanoid platform for research,” in2017 IEEE-RAS 17th International Conference on Humanoid Robotics (Humanoids), 2017, pp. 669–675
2017
-
[16]
Learning to hop for a single-legged robot with parallel mechanism,
H. Zhang, X. Chu, Y . Chen, Y . Tang, L. Yue, Y .-H. Liu, and K. W. S. Au, “Learning to hop for a single-legged robot with parallel mechanism,” arXiv preprint arXiv:2501.11945, 2025
2025 arXiv
-
[17]
Cat-like jumping and landing of legged robots in low gravity using deep reinforcement learning,
N. Rudin, H. Kolvenbach, V . Tsounis, and M. Hutter, “Cat-like jumping and landing of legged robots in low gravity using deep reinforcement learning,”IEEE Transactions on Robotics, vol. 38, no. 1, pp. 317–328, 2022
2022
-
[18]
Booster Gym: An end-to-end reinforcement learning framework for humanoid robot locomotion,
Y . Wang, P. Chen, X. Han, F. Wu, and M. Zhao, “Booster Gym: An end-to-end reinforcement learning framework for humanoid robot locomotion,”arXiv preprint arXiv:2506.15132, 2025
2025 arXiv
-
[19]
Real-world humanoid locomotion with reinforcement learning,
I. Radosavovic, T. Xiao, B. Zhang, T. Darrell, J. Malik, and K. Sreenath, “Real-world humanoid locomotion with reinforcement learning,”Sci- ence Robotics, vol. 9, no. 89, p. eadi9579, 2024
2024
-
[20]
LiPS: Large-scale humanoid robot reinforcement learning with parallel-series structures,
Q. Zhang, G. Han, J. Sun, W. Zhao, J. Cao, J. Wang, H. Cheng, L. Zhang, Y . Guo, and R. Xu, “LiPS: Large-scale humanoid robot reinforcement learning with parallel-series structures,”arXiv preprint arXiv:2503.08349, 2025
2025 arXiv
-
[21]
Control of humanoid robots with parallel mechanisms using differential actuation models,
V . Lutz, L. de Matte ¨ıs, V . Batto, and N. Mansard, “Control of humanoid robots with parallel mechanisms using differential actuation models,” in Proceedings of the 2026 IEEE International Conference on Robotics and Automation (ICRA), Vienna, Austria, Jun. 2026, paper TuI1I.90
2026
-
[22]
Model sim- plification for dynamic control of series–parallel hybrid robots: A representative study on the effects of neglected dynamics,
S. Kumar, J. Martensen, A. Mueller, and F. Kirchner, “Model sim- plification for dynamic control of series–parallel hybrid robots: A representative study on the effects of neglected dynamics,” in2019 IEEE/RSJ International Conference on Intelligent Robots and Systems (IROS), 2...
2019
-
[23]
Sim-to-real transfer of robotic control with dynamics randomization,
X. B. Peng, M. Andrychowicz, W. Zaremba, and P. Abbeel, “Sim-to-real transfer of robotic control with dynamics randomization,” inProceedings of the IEEE International Conference on Robotics and Automation, 2018, pp. 3803–3810
2018
-
[24]
RMA: Rapid motor adaptation for legged robots,
A. Kumar, Z. Fu, D. Pathak, and J. Malik, “RMA: Rapid motor adaptation for legged robots,” inRobotics: Science and Systems, 2021
2021
-
[25]
An efficient learning control framework with sim-to-real for string-type artificial muscle- driven robotic systems,
J. Tao, Y . Zhang, S. K. Rajendran, and F. Zhang, “An efficient learning control framework with sim-to-real for string-type artificial muscle- driven robotic systems,”IEEE/ASME Transactions on Mechatronics, vol. 31, no. 1, pp. 444–455, Feb. 2026
2026
-
[26]
NVIDIA Isaac Sim,
NVIDIA, “NVIDIA Isaac Sim,” Software, version 5.1.0, 2025, released October 30, 2025. [Online]. Available: https://docs.isaacsim.omniverse. nvidia.com/5.1.0/
2025
-
[27]
K. M. Lynch and F. C. Park,Modern Robotics: Mechanics, Planning, and Control. Cambridge University Press, 2017
2017
-
[28]
A unified approach for motion and force control of robot manipulators: The operational space formulation,
O. Khatib, “A unified approach for motion and force control of robot manipulators: The operational space formulation,”IEEE Journal on Robotics and Automation, vol. 3, no. 1, pp. 43–53, Feb. 1987
1987
-
[29]
Synthesis of low-peak-factor signals and binary sequences with low autocorrelation,
M. R. Schroeder, “Synthesis of low-peak-factor signals and binary sequences with low autocorrelation,”IEEE Transactions on Information Theory, vol. 16, no. 1, pp. 85–89, 1970
1970
-
[30]
Pintelon and J
R. Pintelon and J. Schoukens,System Identification: A Frequency Domain Approach, 2nd ed. Wiley-IEEE Press, 2012
2012
-
[31]
Experimental determination of frequency response function estimates for flexible joint industrial manipulators with serial kinematics,
F. Saupe and A. Knoblach, “Experimental determination of frequency response function estimates for flexible joint industrial manipulators with serial kinematics,”Mechanical Systems and Signal Processing, vol. 52– 53, pp. 60–72, 2015
2015
-
[32]
Accurate FRF identification of LPV systems: nD-LPM with application to a medical X-ray system,
R. van der Maas, A. van der Maas, R. J. V oorhoeve, and T. A. E. Oomen, “Accurate FRF identification of LPV systems: nD-LPM with application to a medical X-ray system,”IEEE Transactions on Control Systems Technology, vol. 25, no. 5, pp. 1724–1735, 2017
2017
-
[33]
Prox- imal policy optimization algorithms,
J. Schulman, F. Wolski, P. Dhariwal, A. Radford, and O. Klimov, “Prox- imal policy optimization algorithms,”arXiv preprint arXiv:1707.06347, 2017
2017 arXiv
-
[34]
Sim-to-real: Learning agile locomotion for quadruped robots,
J. Tan, T. Zhang, E. Coumans, A. Iscen, Y . Bai, D. Hafner, S. Bohez, and V . Vanhoucke, “Sim-to-real: Learning agile locomotion for quadruped robots,” inProceedings of Robotics: Science and Systems, 2018
2018
-
[35]
Using data- driven domain randomization to transfer robust control policies to mobile robots,
M. Sheckells, G. Garimella, S. Mishra, and M. Kobilarov, “Using data- driven domain randomization to transfer robust control policies to mobile robots,” in2019 International Conference on Robotics and Automation (ICRA), 2019, pp. 3224–3230
2019
-
[36]
Domain randomization via entropy maximization,
G. Tiboni, P. Klink, J. Peters, T. Tommasi, C. D’Eramo, and G. Chalvatzaki, “Domain randomization via entropy maximization,” in International Conference on Learning Representations, 2024. [Online]. Available: https://openreview.net/forum?id=GXtmuiVrOM
2024
-
[37]
Closing the sim-to-real loop: Adapting simula- tion randomization with real world experience,
Y . Chebotar, A. Handa, V . Makoviychuk, M. Macklin, J. Issac, N. D. Ratliff, and D. Fox, “Closing the sim-to-real loop: Adapting simula- tion randomization with real world experience,” in2019 International Conference on Robotics and Automation (ICRA), 2019, pp. 8973–8979
2019
-
[38]
Mysteric-Net: MIMO hysteretic friction-aware lagrangian-based network for legged robot,
H. Yeo, J. Hong, T. Kong, and S. Oh, “Mysteric-Net: MIMO hysteretic friction-aware lagrangian-based network for legged robot,” in2025 IEEE/RSJ International Conference on Intelligent Robots and Systems (IROS), 2025, pp. 1932–1937
2025
Reviewed August 7, 2026 · model on record in the stance chip above.
Discussion (0). Continue with ORCID to comment.