Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-07-14T11:28:23.511747Z
Paper Citation Record · LEDGER
As of 9 August 2026, this Paper Citation Record lists 85 of 85 outbound references and 0 inbound Pith citation observations for arXiv:2607.10474.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-07-14T11:28:23.511747Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links
A source-named dated measurement, never combined with another source.
Source: cited_works
85 of 85 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 9c2a8297-13ac-4bdb-9160-cda16609d97a · outbound
Reinforcement Learning with Verifiable Physics: Post-training LLMs with Continuous Rewards American Mathematical Society, 2 edition, 2010
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3e264e67-a8dc-4868-a8b1-2be3021a48ab · outbound
Reinforcement Learning with Verifiable Physics: Post-training LLMs with Continuous Rewards Cambridge university press, 2002
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 36fc7b67-b486-4e7c-ac30-93b3ab78b628 · outbound
Reinforcement Learning with Verifiable Physics: Post-training LLMs with Continuous Rewards Springer, 1994
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a5a69d94-9d79-4082-a741-b832f8b37626 · outbound
Reinforcement Learning with Verifiable Physics: Post-training LLMs with Continuous Rewards SIAM, 2000
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 115fd3ce-6941-4823-a23e-6b213aeb88f2 · outbound
Reinforcement Learning with Verifiable Physics: Post-training LLMs with Continuous Rewards SIAM, 1998
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a1a414a4-0922-475a-ab4e-27c0163a690c · outbound
Reinforcement Learning with Verifiable Physics: Post-training LLMs with Continuous Rewards Evaluating Large Language Models Trained on Code
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4b56acf4-4e06-4238-bcb4-8a5b77d36cda · outbound
Reinforcement Learning with Verifiable Physics: Post-training LLMs with Continuous Rewards Measuring Coding Challenge Competence With APPS
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 37cda1a9-6b67-4a43-b293-550fd4b3bf9f · outbound
Reinforcement Learning with Verifiable Physics: Post-training LLMs with Continuous Rewards Program Synthesis with Large Language Models
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cdccca16-61e0-4fb2-a2cc-cb9f75ad6180 · outbound
Reinforcement Learning with Verifiable Physics: Post-training LLMs with Continuous Rewards Coderl: Mastering code generation through pretrained models and deep reinforcement learning
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 130a9fb1-6740-4829-bbfe-fd750d70b16c · outbound
Reinforcement Learning with Verifiable Physics: Post-training LLMs with Continuous Rewards Narasimhan, and Yuan Cao
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 60040e24-02a3-4817-bac6-c7be406df154 · outbound
Reinforcement Learning with Verifiable Physics: Post-training LLMs with Continuous Rewards URLhttps://openreview.net/forum?id=WE_vluYUL-X
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation dfd47062-a949-4b64-969c-22bb38ce2cf8 · outbound
Reinforcement Learning with Verifiable Physics: Post-training LLMs with Continuous Rewards RLEF: Grounding Code LLMs in Execution Feedback with Reinforcement Learning
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation af13fd51-764c-45c3-a680-4bcf05b56b9c · outbound
Reinforcement Learning with Verifiable Physics: Post-training LLMs with Continuous Rewards Codepde: An inference framework for llm-driven PDE solver generation
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7606bdd5-7e76-4301-aae2-79bf3252bb7c · outbound
Reinforcement Learning with Verifiable Physics: Post-training LLMs with Continuous Rewards SciML Agents: Write the Solver, Not the Solution
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d109cd36-cea3-4614-a1b6-1b172d6f578c · outbound
Reinforcement Learning with Verifiable Physics: Post-training LLMs with Continuous Rewards Autonumerics: An autonomous, pde-agnostic multi-agent pipeline for scientific computing.arXiv preprint arXiv:2602.17607, 2026
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 89b9b175-75c9-4af5-8740-28784aeef19d · outbound
Reinforcement Learning with Verifiable Physics: Post-training LLMs with Continuous Rewards All-fem: Agentic large language models fine-tuned for finite element methods.Computer Methods in Applied Mechanics and Engineering, 457:118985, 2026
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 01be39bb-1171-4877-ba82-54d12aebe090 · outbound
Reinforcement Learning with Verifiable Physics: Post-training LLMs with Continuous Rewards Pde-agent: A toolchain-augmented multi-agent framework for pde solving.arXiv preprint arXiv:2512.16214, 2025
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fe4afabf-5ef0-477e-99b5-ba907efa9214 · outbound
Reinforcement Learning with Verifiable Physics: Post-training LLMs with Continuous Rewards PINNsAgent: Automated PDE Surrogation with Large Language Models
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9df63f9d-7f74-4986-b992-8046fc8575f6 · outbound
Reinforcement Learning with Verifiable Physics: Post-training LLMs with Continuous Rewards Scaling LLM Test-Time Compute Optimally can be More Effective than Scaling Model Parameters
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 54306c0d-dc39-4df0-b4bd-4263ab5ae388 · outbound
Reinforcement Learning with Verifiable Physics: Post-training LLMs with Continuous Rewards Inference Scaling Laws: An Empirical Analysis of Compute-Optimal Inference for Problem-Solving with Language Models
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 624f6a03-86d6-4a14-a9db-71e2c052226b · outbound
Reinforcement Learning with Verifiable Physics: Post-training LLMs with Continuous Rewards DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5bbf41bc-d83b-479f-bdd8-71f5097d3cdf · outbound
Reinforcement Learning with Verifiable Physics: Post-training LLMs with Continuous Rewards Unresolved cited work
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 39fc4ebd-d77d-44b6-adb1-6c1a7d4acaf5 · outbound
Reinforcement Learning with Verifiable Physics: Post-training LLMs with Continuous Rewards Solving math word problems with process- and outcome-based feedback
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4d811320-eb90-4cd9-a7f8-44200d21492a · outbound
Reinforcement Learning with Verifiable Physics: Post-training LLMs with Continuous Rewards Efficient memory management for large language model serving with pagedattention
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e84096b0-83c2-4f0e-b77c-6b7afe6c34e1 · outbound
Reinforcement Learning with Verifiable Physics: Post-training LLMs with Continuous Rewards Hybridflow: A flexible and efficient rlhf framework
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 45d4cca4-fa63-48b2-bc1a-8874f9d0d396 · outbound
Reinforcement Learning with Verifiable Physics: Post-training LLMs with Continuous Rewards Huerta, and Hao Peng
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0275feb3-b127-4e03-bd36-874c45a4be81 · outbound
Reinforcement Learning with Verifiable Physics: Post-training LLMs with Continuous Rewards Foam-agent: Towards automated intelligent cfd workflows.arXiv preprint arXiv:2505.04997, 2025
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a85e05d4-686d-4fb3-8def-9ed2414b218d · outbound
Reinforcement Learning with Verifiable Physics: Post-training LLMs with Continuous Rewards Openfoamgpt: A retrieval-augmented large language model (llm) agent for openfoam-based computational fluid dynamics.Physics of Fluids, 37(3), 2025
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1a55b5cc-3704-4caf-94bd-18814d579698 · outbound
Reinforcement Learning with Verifiable Physics: Post-training LLMs with Continuous Rewards MetaOpenFOAM: an LLM-based multi-agent framework for CFD
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2d340ce7-621f-481a-968c-f15309f5d76b · outbound
Reinforcement Learning with Verifiable Physics: Post-training LLMs with Continuous Rewards Mooseagent: A llm based multi-agent framework for automating moose simulation, 2025
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e38baeff-07f0-4af1-800e-ec6d11f182e2 · outbound
Reinforcement Learning with Verifiable Physics: Post-training LLMs with Continuous Rewards Chronollm: customizing language models for physics-based simulation code generation.Multibody System Dynamics, Feb 2026
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 913d1ec4-cdfa-4feb-a229-e36b065e6a8d · outbound
Reinforcement Learning with Verifiable Physics: Post-training LLMs with Continuous Rewards Lang- pinn: From language to physics-informed neural networks via a multi-agent framework.arXiv preprint arXiv:2510.05158, 2025
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fbfc218b-0463-4530-b5e1-d51e4e59d3fd · outbound
Reinforcement Learning with Verifiable Physics: Post-training LLMs with Continuous Rewards FEABench: Evaluating Language Models on Multiphysics Reasoning Ability
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 72c4e00f-6067-4c45-a5a7-e4c55ce47623 · outbound
Reinforcement Learning with Verifiable Physics: Post-training LLMs with Continuous Rewards Solving Physics Olympiad via Reinforcement Learning on Physics Simulators
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1359ee21-87e0-4e85-985e-3ec70721d6e9 · outbound
Reinforcement Learning with Verifiable Physics: Post-training LLMs with Continuous Rewards PDE-Controller: LLMs for Autoformalization and Reasoning of PDEs
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3da208fc-ebdd-4c99-9584-79ae27d54505 · outbound
Reinforcement Learning with Verifiable Physics: Post-training LLMs with Continuous Rewards Agentic scientific simulation: Execution-grounded model construction and reconstruction.arXiv preprint arXiv:2603.00214, 2026
Reference 36
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 84a015bf-0af4-44d6-80c1-f7c88b8bed10 · outbound
Reinforcement Learning with Verifiable Physics: Post-training LLMs with Continuous Rewards DAPO: An open-source LLM reinforcement learning system at scale
Reference 37
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 52864811-cc87-4474-adc2-93ba935dfba9 · outbound
Reinforcement Learning with Verifiable Physics: Post-training LLMs with Continuous Rewards Group Sequence Policy Optimization
Reference 38
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 41ec3765-e6e3-48e6-b353-8aff44293cc7 · outbound
Reinforcement Learning with Verifiable Physics: Post-training LLMs with Continuous Rewards Stepcoder: improving code generation with reinforcement learning from compiler feedback
Reference 39
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 082a7f23-a38a-4140-b7dd-43e98d831d0b · outbound
Reinforcement Learning with Verifiable Physics: Post-training LLMs with Continuous Rewards Reinforcement Learning for Machine Learning Engineering Agents
Reference 40
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f70029e2-a09f-404e-afdb-cc00f2c50e31 · outbound
Reinforcement Learning with Verifiable Physics: Post-training LLMs with Continuous Rewards Beyond Binary Rewards: Training LMs to Reason About Their Uncertainty
Reference 41
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4952e689-6e5e-454c-88ea-06e30e533702 · outbound
Reinforcement Learning with Verifiable Physics: Post-training LLMs with Continuous Rewards Unresolved cited work
Reference 42
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1ec75db1-7a38-471b-b941-8b0b4239a2a6 · outbound
Reinforcement Learning with Verifiable Physics: Post-training LLMs with Continuous Rewards Fourier Neural Operator for Parametric Partial Differential Equations
Reference 43
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 875278d0-a8fd-4ea5-ac02-187f5fd731bf · outbound
Reinforcement Learning with Verifiable Physics: Post-training LLMs with Continuous Rewards Learning nonlinear operators via deeponet based on the universal approximation theorem of operators
Reference 44
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3a8f95d5-ad38-4ff7-b096-65755577ef78 · outbound
Reinforcement Learning with Verifiable Physics: Post-training LLMs with Continuous Rewards Towards long rollout of neural operators with local attention and flow matching-inspired correction: An example in frontal polymerization pdes
Reference 45
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 18612c69-901d-48da-853e-be40c0411fc6 · outbound
Reinforcement Learning with Verifiable Physics: Post-training LLMs with Continuous Rewards Pdebench: An extensive benchmark for scientific machine learning.Advances in neural information processing systems, 35:1596–1611, 2022
Reference 46
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c1e53791-f6a6-4825-9da6-01720d5d4ae0 · outbound
Reinforcement Learning with Verifiable Physics: Post-training LLMs with Continuous Rewards Diffusionpde: Generative pde-solving under partial observation.Advances in Neural Information Processing Systems, 37: 130291–130323, 2024
Reference 47
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 546f6ddc-d397-4231-b52c-47b2f65662d5 · outbound
Reinforcement Learning with Verifiable Physics: Post-training LLMs with Continuous Rewards Physics-Informed Diffusion Models
Reference 48
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0e93baef-bf6f-4707-b999-08b1e1e67819 · outbound
Reinforcement Learning with Verifiable Physics: Post-training LLMs with Continuous Rewards Physics-constrained flow matching: Sampling generative models with hard constraints.arXiv preprint arXiv:2506.04171, 2025
Reference 49
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a1478eb7-a62d-4f7d-bbaf-00b1f33f3939 · outbound
Reinforcement Learning with Verifiable Physics: Post-training LLMs with Continuous Rewards End-to-end probabilistic framework for learning with hard constraints.arXiv preprint arXiv:2506.07003, 2025
Reference 50
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c106f845-3b5f-4fe6-b0d8-551656f32fbe · outbound
Reinforcement Learning with Verifiable Physics: Post-training LLMs with Continuous Rewards Training language models to follow instructions with human feedback
Reference 51
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 62d554c6-900f-468b-9109-ef3307a28e3b · outbound
Reinforcement Learning with Verifiable Physics: Post-training LLMs with Continuous Rewards Qwen2.5-Coder Technical Report
Reference 52
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 25401330-e3f5-4936-804f-83b9a6ad3b99 · outbound
Reinforcement Learning with Verifiable Physics: Post-training LLMs with Continuous Rewards LoRA: Low-Rank Adaptation of Large Language Models
Reference 53
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 18f47ed7-4f88-4f12-8d82-8395aff5f1f8 · outbound
Reinforcement Learning with Verifiable Physics: Post-training LLMs with Continuous Rewards A model for fast computer simulation of waves in excitable media.Physica D: Nonlinear Phenomena, 49(1-2):61–70, 1991
Reference 54
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 29a80986-4a9a-47be-8f06-bdae568ae946 · outbound
Reinforcement Learning with Verifiable Physics: Post-training LLMs with Continuous Rewards Finite difference method for numerical computation of discontinuous solutions of the equations of fluid dynamics.Matematiˇ ceskij sbornik, 47(3): 271–306, 1959
Reference 55
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 262ef592-3736-4460-b7eb-b996676d64bf · outbound
Reinforcement Learning with Verifiable Physics: Post-training LLMs with Continuous Rewards Towards the ultimate conservative difference scheme
Reference 56
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b70b2f42-e862-4aeb-bc88-0fb87d62d5af · outbound
Reinforcement Learning with Verifiable Physics: Post-training LLMs with Continuous Rewards New high-resolution central schemes for nonlinear conservation laws and convection–diffusion equations.Journal of computational physics, 160 (1):241–282, 2000
Reference 57
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation edc2e5e0-c850-4b61-8191-2e4d0b06feda · outbound
Reinforcement Learning with Verifiable Physics: Post-training LLMs with Continuous Rewards Toro.Riemann Solvers and Numerical Methods for Fluid Dynamics
Reference 58
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 47019da4-0458-4c15-9f9a-eb857fe302e9 · outbound
Reinforcement Learning with Verifiable Physics: Post-training LLMs with Continuous Rewards Strong stability-preserving high-order time discretization methods.SIAM review, 43(1):89–112, 2001
Reference 59
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8a4015b7-9d4c-4cd6-b243-f6e25b2ba310 · outbound
Reinforcement Learning with Verifiable Physics: Post-training LLMs with Continuous Rewards On the construction and comparison of difference schemes.SIAM journal on numerical analysis, 5(3):506–517, 1968
Reference 60
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a52ddfc9-a4ff-4487-b268-7fc131ab12d8 · outbound
Reinforcement Learning with Verifiable Physics: Post-training LLMs with Continuous Rewards Implicit-explicit runge-kutta methods for time-dependent partial differential equations.Applied Numerical Mathematics, 25(2-3): 151–167, 1997
Reference 61
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 468f879b-6e28-49e1-811b-476cd6f91fd3 · outbound
Reinforcement Learning with Verifiable Physics: Post-training LLMs with Continuous Rewards Fourth-order time-stepping for stiff pdes.SIAM Journal on Scientific Computing, 26(4):1214–1233, 2005
Reference 62
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a0eeb148-5fcc-417f-8072-fb4fafdb1ba3 · outbound
Reinforcement Learning with Verifiable Physics: Post-training LLMs with Continuous Rewards The numerical solution of the Navier–Stokes equations for an incom- pressible fluid.Bulletin of the American Mathematical Society, 73(6):928–931, 1967
Reference 63
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation c68f4517-646a-46cd-b18f-0f38ebd31974 · outbound
Reinforcement Learning with Verifiable Physics: Post-training LLMs with Continuous Rewards Wellesley-Cambridge Press, 1986
Reference 64
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4c58c301-050a-439f-a395-ad8ceb47c94f · outbound
Reinforcement Learning with Verifiable Physics: Post-training LLMs with Continuous Rewards Scipy 1.0: fundamental algorithms for scientific computing in python.Nature methods, 17(3):261–272, 2020
Reference 65
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ffb451eb-eecf-46eb-b60e-701e506d4255 · outbound
Reinforcement Learning with Verifiable Physics: Post-training LLMs with Continuous Rewards On the elimination of aliasing in finite-difference schemes by filtering high-wavenumber components.Journal of Atmospheric Sciences, 28(6):1074–1074, 1971
Reference 66
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 68bbd1b8-127d-4493-8bbb-b0d3b0c5a434 · outbound
Reinforcement Learning with Verifiable Physics: Post-training LLMs with Continuous Rewards Semi-lagrangian integration schemes for atmospheric models—a review.Monthly weather review, 119(9):2206–2223, 1991
Reference 67
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4458a449-2ca6-4396-9a45-40c8850cd024 · outbound
Reinforcement Learning with Verifiable Physics: Post-training LLMs with Continuous Rewards SIAM, 2007
Reference 68
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cccf25a3-36ab-4c17-9633-25e48b97dfd3 · outbound
Reinforcement Learning with Verifiable Physics: Post-training LLMs with Continuous Rewards The calculation of the interaction of non-stationary shock waves and obstacles.USSR Computational Mathematics and Mathematical Physics, 1(2):304–320, 1962
Reference 69
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2826c273-9ea2-42ed-b809-f5e36decb625 · outbound
Reinforcement Learning with Verifiable Physics: Post-training LLMs with Continuous Rewards Systems of conservation laws.Communications on Pure and Applied Mathematics, 13:217–237, 1960
Reference 70
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f3bceee2-ffed-4c61-98dc-29d9215b16fb · outbound
Reinforcement Learning with Verifiable Physics: Post-training LLMs with Continuous Rewards The effect of viscosity in hypervelocity impact cratering.Journal of spacecraft and rockets, 40(5):757–763, 2003
Reference 71
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation df339ee1-c52c-457b-bd41-89605c43c313 · outbound
Reinforcement Learning with Verifiable Physics: Post-training LLMs with Continuous Rewards Springer, 2003
Reference 72
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d93d8694-048c-474b-b93d-9c9e2cdedd53 · outbound
Reinforcement Learning with Verifiable Physics: Post-training LLMs with Continuous Rewards Finite volume methods.Handbook of numerical analysis, 7:713–1018, 2000
Reference 73
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cdda3a8f-19cc-4c8d-b1da-891fd4a6ac2e · outbound
Reinforcement Learning with Verifiable Physics: Post-training LLMs with Continuous Rewards Methods of conjugate gradients for solving linear systems.Journal of research of the National Bureau of Standards, 49(6):409–436, 1952
Reference 74
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ef4648b6-41a3-46da-8d2e-b9292b759f08 · outbound
Reinforcement Learning with Verifiable Physics: Post-training LLMs with Continuous Rewards SIAM, 2003
Reference 75
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f246ce2d-75ae-44a4-b0d5-9da9de767a50 · outbound
Reinforcement Learning with Verifiable Physics: Post-training LLMs with Continuous Rewards Springer, 4 edition, 2020
Reference 76
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3ae2cb35-29a7-43c5-9805-e18c08913990 · outbound
Reinforcement Learning with Verifiable Physics: Post-training LLMs with Continuous Rewards Approximate riemann solvers, parameter vectors, and difference schemes.Journal of computational physics, 43(2):357–372, 1981
Reference 77
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation eb47f0e1-c1c6-4c64-b552-c5e5c29897be · outbound
Reinforcement Learning with Verifiable Physics: Post-training LLMs with Continuous Rewards Information theory and statistical mechanics.Physical review, 106(4):620, 1957
Reference 78
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 48dbb034-9c96-4bb9-b7f8-c22117313ee6 · outbound
Reinforcement Learning with Verifiable Physics: Post-training LLMs with Continuous Rewards Academic Press, 11 edition, 2014
Reference 79
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 55b2cfab-85c7-4f2e-ac09-efe89111e1e0 · outbound
Reinforcement Learning with Verifiable Physics: Post-training LLMs with Continuous Rewards Equation of state calculations by fast computing machines.The journal of chemical physics, 21(6):1087–1092, 1953
Reference 80
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8877ab60-e240-4a88-82e8-1fbcb8052472 · outbound
Reinforcement Learning with Verifiable Physics: Post-training LLMs with Continuous Rewards Ziebart, Andrew Maas, J
Reference 81
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8cf82fa4-923f-4b76-b026-9384fe8b1903 · outbound
Reinforcement Learning with Verifiable Physics: Post-training LLMs with Continuous Rewards Soft actor-critic: Off- policy maximum entropy deep reinforcement learning with a stochastic actor
Reference 82
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6d26fc3b-1ddb-4815-8dc5-f8858144caac · outbound
Reinforcement Learning with Verifiable Physics: Post-training LLMs with Continuous Rewards Where applicable, we additionally require that solvers within a scheme family agree on smooth initial data to within their formal order at fixed resolution
Reference 83
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 499fcfd1-cbfd-4bbf-99c1-f1e06c0457ae · outbound
Reinforcement Learning with Verifiable Physics: Post-training LLMs with Continuous Rewards This producesabsoluteerror against the exact ue at each grid level rather than a self-consistency ratio, and detects sign errors and stencil bugs that self- convergence cannot
Reference 84
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c849d7e5-efa8-43c8-9412-8a447c9a0809 · outbound
Reinforcement Learning with Verifiable Physics: Post-training LLMs with Continuous Rewards This provides a check against an independent benchmark, complementing self-convergence and MMS
Reference 85
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
No inbound Pith citation observations are available.