Pith. sign in

Paper Citation Record · LEDGER

AnswerCarefully: A Dataset for Improving the Safety of Japanese LLM Output

As of 18 August 2026, this Paper Citation Record lists 18 of 18 outbound references and 1 inbound Pith citation observation for arXiv:2506.02372.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2506.02372 v1

Coverage vector

measured 18 of 18 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T11:29:15.193096Z

measured 19 of 19 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-17T06:30:58.91139+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T10:17:26.850823Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

18 of 18 outbound references displayed

  • verified exact0
  • verified fuzzy9
  • unresolved9
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 258bae2d-697f-433d-a09f-efaae2cac962 · outbound

This paper cites Training a Helpful and Harmless Assistant with Reinforcement Learning from Human Feedback.

AnswerCarefully: A Dataset for Improving the Safety of Japanese LLM Output Training a Helpful and Harmless Assistant with Reinforcement Learning from Human Feedback

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-07T11:29:13.172924Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:29:13.172924Z digest=sha256:8ce78c5f90ab2df233ff47ee5975b4acddb43269d72083ffdf3f3932f2ae8565

Observation 5d43c3a6-5a43-45d0-b0c8-a0efb2f51d51 · outbound

This paper cites Bowman, Zac Hatfield-Dodds, Ben Mann, Dario Amodei, Nicholas Joseph, Sam McCandlish, Tom Brown, and Jared Kaplan.

AnswerCarefully: A Dataset for Improving the Safety of Japanese LLM Output Bowman, Zac Hatfield-Dodds, Ben Mann, Dario Amodei, Nicholas Joseph, Sam McCandlish, Tom Brown, and Jared Kaplan

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:29:17.996876Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T11:29:13.228906Z digest=sha256:872fe7593d520f404cd473d18e37d65ac7ebeffd6f0aee08658a51508161af33

Observation 312bed32-3c6c-4f5a-a6d9-bba5dc19d8e0 · outbound

This paper cites Llama Guard: LLM-based Input-Output Safeguard for Human-AI Conversations.

AnswerCarefully: A Dataset for Improving the Safety of Japanese LLM Output Llama Guard: LLM-based Input-Output Safeguard for Human-AI Conversations

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-07T11:29:13.320214Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:29:13.320214Z digest=sha256:3536ac0376c9883413a5befb105d7e190b5aba479c4f06aa4a0551520794156e

Observation 1bec56d0-8f72-4102-8c4a-a814d710fabb · outbound

This paper cites Investigating tuning methods for achieving usefulness and safety for Japanese large language models (in Japanese).

AnswerCarefully: A Dataset for Improving the Safety of Japanese LLM Output Investigating tuning methods for achieving usefulness and safety for Japanese large language models (in Japanese)

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:29:17.723376Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T11:29:13.460840Z digest=sha256:47f56faf9b8eee21f2c5e091a1f01e686f98f677f327d7dc05e4b857542800c6

Observation e95eafd8-5227-4877-85f6-b62b11c09b86 · outbound

This paper cites Japanese safety boundary test for large language models (in Japanese).

AnswerCarefully: A Dataset for Improving the Safety of Japanese LLM Output Japanese safety boundary test for large language models (in Japanese)

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:29:17.503783Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T11:29:13.572380Z digest=sha256:4fea2a8878cfec7adefccd2ea9b3baa4bf6d4f0961f268592f1a9224da5ac6b6

Observation 0b3b0dce-8bbd-4034-baf7-21f25fdb39a8 · outbound

This paper cites LLM-jp: A Cross-organizational Project for the Research and Development of Fully Open Japanese LLMs.

AnswerCarefully: A Dataset for Improving the Safety of Japanese LLM Output LLM-jp: A Cross-organizational Project for the Research and Development of Fully Open Japanese LLMs

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-07T11:29:13.679486Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:29:13.679486Z digest=sha256:8a0e28e351c1aa452e35513184bfe9bbf236e1344766e8607706d16d72882c32

Observation cfe4d581-d78d-43c8-b876-8dfb84542e94 · outbound

This paper cites Construction of the Japanese TruthfulQA Dataset (in Japanese).

AnswerCarefully: A Dataset for Improving the Safety of Japanese LLM Output Construction of the Japanese TruthfulQA Dataset (in Japanese)

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:29:17.253806Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T11:29:13.781769Z digest=sha256:a6a45d7b82e132b662e4025c95486185a74b0968ced9187af68ed226eb181145

Observation 7c5ba52e-7c82-4aa3-8f24-a59be2b46736 · outbound

This paper cites an unresolved cited work.

AnswerCarefully: A Dataset for Improving the Safety of Japanese LLM Output Unresolved cited work

Reference 8

Resolution
unresolved
raw_fallback, observed 2026-08-07T11:29:17.034182Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T11:29:13.929012Z digest=sha256:d877572edfa4fa549ac05010d847f0ad036b078da645ceb8682ab9ce2849680f

Observation 8e5070fc-ede0-47ef-968b-62f442e5007f · outbound

This paper cites JSocialFact: a misinfor- mation dataset from social media for benchmarking LLM safety.

AnswerCarefully: A Dataset for Improving the Safety of Japanese LLM Output JSocialFact: a misinfor- mation dataset from social media for benchmarking LLM safety

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-07T11:29:14.087472Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:29:14.087472Z digest=sha256:ff48c288d7d0f983929d074945b2192aa7f9055e5a086bf195e655f5a6004f8e

Observation 50ad9775-4563-4de8-acd6-317acfea1bab · outbound

This paper cites GPT-4 technical report, 2024.

AnswerCarefully: A Dataset for Improving the Safety of Japanese LLM Output GPT-4 technical report, 2024

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:29:16.755003Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T11:29:14.219207Z digest=sha256:fc7906562a03127ffc308e42bd5fd2c8ad3fdcf56347223d063828da7f8ecd00

Observation e171592d-8d21-4485-b4d4-09aacb7b3792 · outbound

This paper cites Large-scale human evaluation of LLM safety (in Japanese).

AnswerCarefully: A Dataset for Improving the Safety of Japanese LLM Output Large-scale human evaluation of LLM safety (in Japanese)

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:29:16.526108Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T11:29:14.329033Z digest=sha256:c96639ebc90864ef713fcd2b1a1f3be6cb039d22d0ea6d376537fced59141cf2

Observation be382a7c-2b54-43c8-ac68-79a9c93bbd81 · outbound

This paper cites Gemini: A family of highly capable multimodal models, 2024.

AnswerCarefully: A Dataset for Improving the Safety of Japanese LLM Output Gemini: A family of highly capable multimodal models, 2024

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-07T11:29:14.456053Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:29:14.456053Z digest=sha256:f266b1268fcc1ea2541841a719b9944c264c222adbba1128438024ba0b7d0685

Observation da3c5d52-8b6b-47f7-b631-89c647c403d6 · outbound

This paper cites Llama 2: Open foundation and fine-tuned chat models, 2023.

AnswerCarefully: A Dataset for Improving the Safety of Japanese LLM Output Llama 2: Open foundation and fine-tuned chat models, 2023

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:29:16.303653Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T11:29:14.562203Z digest=sha256:c3019b5f9640ffbd525372d2d4c912ffc48bd23a6440b7f18c15fbb81c945222

Observation 36eafbe0-8e45-4677-b670-6bfbb183a6df · outbound

This paper cites Do-Not-Answer: Evaluating safeguards in LLMs.

AnswerCarefully: A Dataset for Improving the Safety of Japanese LLM Output Do-Not-Answer: Evaluating safeguards in LLMs

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:29:16.035331Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T11:29:14.670540Z digest=sha256:22ab16a951fbcb730f0015225467032ad66bbbe82b09bd03b337e64099772b92

Observation 27c19c1a-5550-4496-9d84-6c4786339b2c · outbound

This paper cites A Chinese Dataset for Evaluating the Safeguards in Large Language Models.

AnswerCarefully: A Dataset for Improving the Safety of Japanese LLM Output A Chinese Dataset for Evaluating the Safeguards in Large Language Models

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-07T11:29:14.775471Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:29:14.775471Z digest=sha256:1d4ad7f4e30b6f38de57ab23e95cf0a9c45f0ce821717c917a1f0441a5524db1

Observation 6bff4682-60b8-46d1-b28a-b4bfe646b41c · outbound

This paper cites JBBQ: Japanese Bias Benchmark for Analyzing Social Biases in Large Language Models.

AnswerCarefully: A Dataset for Improving the Safety of Japanese LLM Output JBBQ: Japanese Bias Benchmark for Analyzing Social Biases in Large Language Models

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-07T11:29:15.043656Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:29:15.043656Z digest=sha256:9d92063a5122c643532f9cb2c117daff3a43d6bf22eab2f56ec89b457f554d81

Observation 76600452-782e-4b5d-94a5-c31bbd0e8567 · outbound

This paper cites Xing, Hao Zhang, Joseph E.

AnswerCarefully: A Dataset for Improving the Safety of Japanese LLM Output Xing, Hao Zhang, Joseph E

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:29:15.603317Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T11:29:15.193096Z digest=sha256:8facc3a7198736a8ceeae4de72600919dd779d4d6d7a14ea058e052f1028139a

Observation b6abdcf6-0478-4e02-806e-adfe4b6e71f7 · outbound

This paper cites an unresolved cited work.

AnswerCarefully: A Dataset for Improving the Safety of Japanese LLM Output Unresolved cited work

Reference 184

Resolution
unresolved
raw_fallback, observed 2026-08-07T11:29:15.828088Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T11:29:14.906067Z digest=sha256:3a5f3915854f82b373ece0ef0acd07f47b3bd599adfb851f0ec807efd68b1fb4

Pith citing papers

Observation 5cf06d1a-4497-4622-81af-3c946d84e8ef · inbound

The Scales of Justitia: A Comprehensive Survey on Safety Evaluation of LLMs cites this paper.

The Scales of Justitia: A Comprehensive Survey on Safety Evaluation of LLMs AnswerCarefully: A Dataset for Improving the Safety of Japanese LLM Output

Reference 85

Resolution
unresolved
no resolver link, observed 2026-08-07T10:17:26.850823Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T10:17:26.850823Z digest=sha256:6b8b8aeede2d81786375209d5af75f99981ee1c339ed0b9996419c2f1787cf06