Pith. sign in

Paper Citation Record · LEDGER

Zeroth-Order Optimization Meets Human Feedback: Provable Learning via Ranking Oracles

As of 12 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 8 inbound Pith citation observations for arXiv:2303.03751.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2303.03751 v3

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 8 of 8 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-12T06:34:41.77262+00:00

measured 8 of 8 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-12T13:55:43.053879Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-04T13:09:51.325412Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation c11513d6-9d16-407a-b038-04becfbf1ecd · inbound

Ruppert-Polyak averaging for Stochastic Order Oracle cites this paper.

Ruppert-Polyak averaging for Stochastic Order Oracle Zeroth-Order Optimization Meets Human Feedback: Provable Learning via Ranking Oracles

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-12T13:55:43.053879Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T13:55:43.053879Z digest=sha256:0d3aa44676a85cf77f218fc80628adf33d81aa377a9da4323dbdf8d46136eda7

Observation 65461040-35cb-4b06-8965-528cc7ed8e69 · inbound

Hierarchical Prompt Decision Transformer: Improving Few-Shot Policy Generalization with Global and Adaptive Guidance cites this paper.

Hierarchical Prompt Decision Transformer: Improving Few-Shot Policy Generalization with Global and Adaptive Guidance Zeroth-Order Optimization Meets Human Feedback: Provable Learning via Ranking Oracles

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-12T04:52:59.579013Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T04:52:59.579013Z digest=sha256:3d256c8b58c8b0ef5c414456f68c636419e81a521571fd2d60b18ef235e38cf0

Observation 6e240516-5ea6-4089-99c3-efbd86510ea0 · inbound

Prompt-Tuning Bandits: Enabling Few-Shot Generalization for Efficient Multi-Task Offline RL cites this paper.

Prompt-Tuning Bandits: Enabling Few-Shot Generalization for Efficient Multi-Task Offline RL Zeroth-Order Optimization Meets Human Feedback: Provable Learning via Ranking Oracles

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-08T15:45:06.712104Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T15:45:06.712104Z digest=sha256:5b7e42b2bfc0958fe4a7b0943048005eaf6a976be64287fb021c1ab2795bca24

Observation d54c479e-6732-4401-8172-542ba2789ed3 · inbound

Instant Preference Alignment for Text-to-Image Diffusion Models cites this paper.

Instant Preference Alignment for Text-to-Image Diffusion Models Zeroth-Order Optimization Meets Human Feedback: Provable Learning via Ranking Oracles

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-05T16:51:04.814963Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T16:51:04.814963Z digest=sha256:5e13e2fd3f3dfcac78799b5f646d48ada59890639cc7f9575709fdaf608e60f3

Observation 3c35af9b-7a2e-4d5e-abd8-1961b0bcda92 · inbound

Oracle Noise: Faster Semantic Spherical Alignment for Interpretable Latent Optimization cites this paper.

Oracle Noise: Faster Semantic Spherical Alignment for Interpretable Latent Optimization Zeroth-Order Optimization Meets Human Feedback: Provable Learning via Ranking Oracles

Reference 39

Resolution
verified exact
arxiv_id, observed 2026-05-11T21:06:12.114765Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-05-08T06:58:09.085573Z digest=sha256:d447948a0effc9060768ae1017893ebb407cdd74667db6e329cb5abf69092102

Observation 61092903-4699-4fd0-96ba-ad0feaca4269 · inbound

AdaMeZO: Adam-style Zeroth-Order Optimizer for LLM Fine-tuning Without Maintaining the Moments cites this paper.

AdaMeZO: Adam-style Zeroth-Order Optimizer for LLM Fine-tuning Without Maintaining the Moments Zeroth-Order Optimization Meets Human Feedback: Provable Learning via Ranking Oracles

Reference 64

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T15:31:07.649291Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-05-09T19:50:50.653184Z digest=sha256:7f9740e11726d96765ab8280c523d526c70e36ae7ec23664b784d7fead22f754

Observation 34afb4dd-5ed6-4a44-a72c-c89865faee4c · inbound

Finding Stationary Points by Comparisons cites this paper.

Finding Stationary Points by Comparisons Zeroth-Order Optimization Meets Human Feedback: Provable Learning via Ranking Oracles

Reference 43

Resolution
metadata mismatch
arxiv_id, observed 2026-07-04T13:09:51.326942Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-06-26T05:26:07.720218Z digest=sha256:39a8ffe33b87b596935966d0c01677dcb996ecbd30cfa9a476a5d9aa5ee28154

Observation 599084a4-c3c0-4de6-99da-26acea2494d4 · inbound

Sculpting NeRF Geometry: Human-Preference Fine-Tuning of a 3D-Aware Face GAN cites this paper.

Sculpting NeRF Geometry: Human-Preference Fine-Tuning of a 3D-Aware Face GAN Zeroth-Order Optimization Meets Human Feedback: Provable Learning via Ranking Oracles

Reference 56

Resolution
verified exact
arxiv_id, observed 2026-07-04T12:59:53.187149Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-06-26T05:33:57.816715Z digest=sha256:76f1e7880b0832d46e6fff30bb9a9ac827e2aa75d887dcbd3035886c0cf56186