Pith. sign in

Paper Citation Record · LEDGER

From Interpretability to Control: Insights from Six Years of the TrustNLP Workshop

As of 19 August 2026, this Paper Citation Record lists 100 of 118 outbound references and 0 inbound Pith citation observations for arXiv:2608.11171.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2608.11171 v1

Coverage vector

measured 100 of 118 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-12T04:51:17.259117Z

measured 100 of 100 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-19T06:32:44.657259+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

100 of 118 outbound references displayed

  • verified exact6
  • verified fuzzy12
  • unresolved82
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 8d9ca6f8-ec40-4cd2-9a0e-f359d768f7d3 · outbound

This paper cites an unresolved cited work.

From Interpretability to Control: Insights from Six Years of the TrustNLP Workshop Unresolved cited work

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-12T04:51:16.630599Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T04:51:16.630599Z digest=sha256:acca948f3d07680657f5891bc29f1fdfbf7546bda1f3cacdcd7f029311d7fb65

Observation de088d98-e2f9-4e2d-af73-5700f1053308 · outbound

This paper cites an unresolved cited work.

From Interpretability to Control: Insights from Six Years of the TrustNLP Workshop Unresolved cited work

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-12T04:51:16.648681Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T04:51:16.648681Z digest=sha256:84f2eeeb837b2fe8d49fd16506148cfce9807166fcb470db3f9df18c3d1e55e7

Observation bf5ef2da-7092-4ecf-aada-16d42922fd86 · outbound

This paper cites an unresolved cited work.

From Interpretability to Control: Insights from Six Years of the TrustNLP Workshop Unresolved cited work

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-12T04:51:16.667018Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T04:51:16.667018Z digest=sha256:b8cb7432ffe29c6666522db5e628a9f788835328f488c6091abad0d67064dde4

Observation c011e2b4-367e-4497-9fa6-fa23c58f332f · outbound

This paper cites an unresolved cited work.

From Interpretability to Control: Insights from Six Years of the TrustNLP Workshop Unresolved cited work

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-12T04:51:16.671317Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T04:51:16.671317Z digest=sha256:c41a761a8b798fedc7f37143eb63d3b619bfe875433bc650fb0170999455ddc0

Observation c8e8bca1-0a47-4526-81e7-a4ff31fb9f57 · outbound

This paper cites an unresolved cited work.

From Interpretability to Control: Insights from Six Years of the TrustNLP Workshop Unresolved cited work

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-12T04:51:16.685034Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T04:51:16.685034Z digest=sha256:b4303f6e915b9964141e2afeb329db81655a0667ee76d2dbf57f4a48fd753b1f

Observation baac49ac-a54a-4be1-b053-273459137e61 · outbound

This paper cites an unresolved cited work.

From Interpretability to Control: Insights from Six Years of the TrustNLP Workshop Unresolved cited work

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-12T04:51:16.690286Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T04:51:16.690286Z digest=sha256:556fd43e79499441fdfaa55de9070bdb5a8455d2d753a9e13cbb468255b312d6

Observation c4ffb726-c5f2-42d2-b97f-597a473e3724 · outbound

This paper cites don't forget the teachers.

From Interpretability to Control: Insights from Six Years of the TrustNLP Workshop don't forget the teachers

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-12T04:51:16.737630Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T04:51:16.737630Z digest=sha256:c8ce9091496371b5b8ae907aecf8f204c5a39922ca89c1ffd9961ab4702e6b8e

Observation 6b16f8c3-3af4-417c-bddc-72aef843c6b5 · outbound

This paper cites an unresolved cited work.

From Interpretability to Control: Insights from Six Years of the TrustNLP Workshop Unresolved cited work

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-12T04:51:16.745842Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T04:51:16.745842Z digest=sha256:07d423dc5481799976ecb9e26977a229bd13054b23168cd913a5af5e1131bbf7

Observation 9870e611-b528-4ab9-9e31-792d137b9324 · outbound

This paper cites TrustLLM: Trustworthiness in Large Language Models.

From Interpretability to Control: Insights from Six Years of the TrustNLP Workshop TrustLLM: Trustworthiness in Large Language Models

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-12T04:51:16.754194Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T04:51:16.754194Z digest=sha256:c6f9ddada40dc97f913421bd6b15ce9481ecc664d3c2333c07ecb24468dd6442

Observation 46401178-93bd-4429-a47e-4e174c89416b · outbound

This paper cites an unresolved cited work.

From Interpretability to Control: Insights from Six Years of the TrustNLP Workshop Unresolved cited work

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-12T04:51:16.758736Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T04:51:16.758736Z digest=sha256:7e08452a6f3ebe25ef91af542b35a6a3f226ed2e997d3c1ff36f93b607f91882

Observation 831c3195-f441-40e5-a3e8-22ca521a3feb · outbound

This paper cites an unresolved cited work.

From Interpretability to Control: Insights from Six Years of the TrustNLP Workshop Unresolved cited work

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-12T04:51:16.767502Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T04:51:16.767502Z digest=sha256:b4470f25c6d4c418f2d75adda879ed8804af717f4b68a5f2fc610ae0cf4c3269

Observation 163ffdfc-8028-4fbd-a4a9-71c18d0b6783 · outbound

This paper cites an unresolved cited work.

From Interpretability to Control: Insights from Six Years of the TrustNLP Workshop Unresolved cited work

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-12T04:51:16.771478Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T04:51:16.771478Z digest=sha256:87fc86789f68456bcaff46e380bf9642de739434ed5c3576f680d11d89cc919c

Observation 1a998e40-6912-4942-baef-7f9c23cbe17e · outbound

This paper cites an unresolved cited work.

From Interpretability to Control: Insights from Six Years of the TrustNLP Workshop Unresolved cited work

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-12T04:51:16.775652Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T04:51:16.775652Z digest=sha256:59b0f409f1337119035057b5831b1d543b6cd274c4dae985c89d7f99c15a8de0

Observation 645cca4e-9f50-4fb8-93ce-98e5cd658204 · outbound

This paper cites an unresolved cited work.

From Interpretability to Control: Insights from Six Years of the TrustNLP Workshop Unresolved cited work

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-12T04:51:16.811315Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T04:51:16.811315Z digest=sha256:66fcc6cf80082a65e4265913c316526daf1e00cfc0021706efd3885a0650f2ea

Observation f74c9f77-3b20-4494-981d-04f072dc2f58 · outbound

This paper cites an unresolved cited work.

From Interpretability to Control: Insights from Six Years of the TrustNLP Workshop Unresolved cited work

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-12T04:51:16.815380Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T04:51:16.815380Z digest=sha256:cab0c6e074a934f2343540cb215f5168301d30ee4faf53fba17e1bdde65e3276

Observation b1e16acf-8f1e-42f5-a2a3-a5e70d8948fa · outbound

This paper cites an unresolved cited work.

From Interpretability to Control: Insights from Six Years of the TrustNLP Workshop Unresolved cited work

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-12T04:51:16.827962Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T04:51:16.827962Z digest=sha256:149ad85bef8fc04048cf96d129491509efd988997100781405951d8a2c143a34

Observation 6041a784-e8cb-42c8-aa8a-ad32cb84e29a · outbound

This paper cites an unresolved cited work.

From Interpretability to Control: Insights from Six Years of the TrustNLP Workshop Unresolved cited work

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-12T04:51:16.857656Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T04:51:16.857656Z digest=sha256:da43e27448b3408da75ba79784c00445cc7073a1e1e1cb6c57e026ef366b1e82

Observation 7cddb846-25b5-4780-9965-db36f7c79c4b · outbound

This paper cites an unresolved cited work.

From Interpretability to Control: Insights from Six Years of the TrustNLP Workshop Unresolved cited work

Reference 59

Resolution
unresolved
no resolver link, observed 2026-08-12T04:51:16.880222Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T04:51:16.880222Z digest=sha256:5337afcb789c553673ef8b5d8b664592de18368c216254ada0e1ad568bff46ce

Observation 68a32cee-7b52-4259-8f3a-8fe60d1019e6 · outbound

This paper cites DecodingTrust: A Comprehensive Assessment of Trustworthiness in GPT Models.

From Interpretability to Control: Insights from Six Years of the TrustNLP Workshop DecodingTrust: A Comprehensive Assessment of Trustworthiness in GPT Models

Reference 61

Resolution
unresolved
no resolver link, observed 2026-08-12T04:51:16.888702Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T04:51:16.888702Z digest=sha256:1ff4f77547bc2d3ef09627ca669fa8bc2c714e65283c4acaca693849320e1de4

Observation 02bafdd9-53fb-4a9a-a152-1238249fb910 · outbound

This paper cites an unresolved cited work.

From Interpretability to Control: Insights from Six Years of the TrustNLP Workshop Unresolved cited work

Reference 64

Resolution
unresolved
no resolver link, observed 2026-08-12T04:51:16.901804Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T04:51:16.901804Z digest=sha256:65f452939e5deee0d1476bc4d55993667c3505605ef0e98f4e4cc507d61cf3be

Observation 3c7f868c-804e-474f-8090-2f9f3b95bf1c · outbound

This paper cites an unresolved cited work.

From Interpretability to Control: Insights from Six Years of the TrustNLP Workshop Unresolved cited work

Reference 68

Resolution
unresolved
no resolver link, observed 2026-08-12T04:51:16.919064Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T04:51:16.919064Z digest=sha256:0e2572e53722978e1525e0c8e6d51c1b5128105dc076db3a9afc1060c0b8b14f

Observation 0c8f1fc8-d6fe-4850-8274-79fed12bacdf · outbound

This paper cites an unresolved cited work.

From Interpretability to Control: Insights from Six Years of the TrustNLP Workshop Unresolved cited work

Reference 69

Resolution
unresolved
no resolver link, observed 2026-08-12T04:51:16.923408Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T04:51:16.923408Z digest=sha256:91cd50e66bab2b53312664c6aa1f642375e7331f21a7f7c046de5cbc9270f9ee

Observation db943417-ff45-4671-91b0-ed7794a7416a · outbound

This paper cites an unresolved cited work.

From Interpretability to Control: Insights from Six Years of the TrustNLP Workshop Unresolved cited work

Reference 70

Resolution
unresolved
no resolver link, observed 2026-08-12T04:51:16.927400Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T04:51:16.927400Z digest=sha256:6e8d2c8eac55f41ad8739a78a6397d7fca05d8a9209dbc7492893c306432f089

Observation c3b4fb2d-528f-4e0a-89e8-0382adcebc0b · outbound

This paper cites an unresolved cited work.

From Interpretability to Control: Insights from Six Years of the TrustNLP Workshop Unresolved cited work

Reference 71

Resolution
unresolved
no resolver link, observed 2026-08-12T04:51:16.931692Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T04:51:16.931692Z digest=sha256:e5f5b679819e83b32a9604a2ae3835b1d0dd6eb45c039137c107afd5bceab79f

Observation c479cfcd-9596-4125-8055-ad17f421356b · outbound

This paper cites an unresolved cited work.

From Interpretability to Control: Insights from Six Years of the TrustNLP Workshop Unresolved cited work

Reference 72

Resolution
unresolved
no resolver link, observed 2026-08-12T04:51:16.935849Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T04:51:16.935849Z digest=sha256:9c96862e5518b0e7b31338892b026ae1b7b38173bc4c0b9a3fe6e76e10f36869

Observation 84c9f417-3d2e-4107-bed1-25062866139e · outbound

This paper cites 2026 , howpublished =.

From Interpretability to Control: Insights from Six Years of the TrustNLP Workshop 2026 , howpublished =

Reference 73

Resolution
unresolved
no resolver link, observed 2026-08-12T04:51:16.939878Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T04:51:16.939878Z digest=sha256:aae2cac0e1c00d40730ffa04f9ab98998105a850ddc9c144bb47620f02a95c9f

Observation 26f9553c-e3d4-432a-a788-02e54cf8dc1e · outbound

This paper cites Human-Centered Explainable AI : Towards a Reflective Sociotechnical Approach.

From Interpretability to Control: Insights from Six Years of the TrustNLP Workshop Human-Centered Explainable AI : Towards a Reflective Sociotechnical Approach

Reference 74

Resolution
unresolved
no resolver link, observed 2026-08-12T04:51:16.944004Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T04:51:16.944004Z digest=sha256:4919828a19b5aeec62107be8d4a9803c23937d323d6d26e67cdfc27f8c5cc663

Observation 107975d5-4d78-4da4-8757-05817c127b83 · outbound

This paper cites and Wintersberger, Philipp and Manger, Carina and Hubig, Nina and Savage, Saiph and Weisz, Justin D.

From Interpretability to Control: Insights from Six Years of the TrustNLP Workshop and Wintersberger, Philipp and Manger, Carina and Hubig, Nina and Savage, Saiph and Weisz, Justin D

Reference 75

Resolution
unresolved
no resolver link, observed 2026-08-12T04:51:16.948539Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T04:51:16.948539Z digest=sha256:6cec25980c96fee0c737c7da36b16cedc359666f7e0ee5fa7c8c9d2dc47bd508

Observation 0e2ea7ae-a4f9-46c1-9066-b8b0cf35c35a · outbound

This paper cites A Survey on Medical Large Language Models: Technology, Application, Trustworthiness, and Future Directions.

From Interpretability to Control: Insights from Six Years of the TrustNLP Workshop A Survey on Medical Large Language Models: Technology, Application, Trustworthiness, and Future Directions

Reference 76

Resolution
unresolved
no resolver link, observed 2026-08-12T04:51:16.952670Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T04:51:16.952670Z digest=sha256:eeba3997ce987ca6db4b3c7214a3b7c524a38f9b96fdeb62196465554f00609e

Observation da851883-7758-46cc-820b-29c1860d8b97 · outbound

This paper cites Standard Benchmarks Fail -- Auditing LLM Agents in Finance Must Prioritize Risk.

From Interpretability to Control: Insights from Six Years of the TrustNLP Workshop Standard Benchmarks Fail -- Auditing LLM Agents in Finance Must Prioritize Risk

Reference 77

Resolution
unresolved
no resolver link, observed 2026-08-12T04:51:16.956798Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T04:51:16.956798Z digest=sha256:34fa24f8ffb90f336bb2eee432f5de8bb86bde494c14c3d70e282b7719415c67

Observation 91bc9c9d-60ba-4cbf-9cf8-667c86e04d82 · outbound

This paper cites Don't Forget the Teachers.

From Interpretability to Control: Insights from Six Years of the TrustNLP Workshop Don't Forget the Teachers

Reference 78

Resolution
unresolved
no resolver link, observed 2026-08-12T04:51:16.960742Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T04:51:16.960742Z digest=sha256:62f5c221f4fa9e056b9944dc03fdbe47ebe86b4e4a5258f6d8f9ae17837cb0ec

Observation 913fa84b-6be8-4080-8c9f-6d436179cc43 · outbound

This paper cites 2025 Silicon Valley Cybersecurity Conference (SVCC) , pages=.

From Interpretability to Control: Insights from Six Years of the TrustNLP Workshop 2025 Silicon Valley Cybersecurity Conference (SVCC) , pages=

Reference 79

Resolution
unresolved
no resolver link, observed 2026-08-12T04:51:16.965221Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T04:51:16.965221Z digest=sha256:1fd77696ee7a6dd8c91532ebb9744188c5aeb77e8c237476f8d57abae1a1483e

Observation 3bd9c0e2-d80d-419f-a421-d48a1d9b36af · outbound

This paper cites 2023 , url =.

From Interpretability to Control: Insights from Six Years of the TrustNLP Workshop 2023 , url =

Reference 80

Resolution
unresolved
no resolver link, observed 2026-08-12T04:51:16.969245Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T04:51:16.969245Z digest=sha256:6e35d8db54e84df55e077ba0afa58bb58792b9c18e44a798540ec5b6158f42a3

Observation 4cf52372-36ab-4930-b5d2-9427461bb01c · outbound

This paper cites 2024 , url=.

From Interpretability to Control: Insights from Six Years of the TrustNLP Workshop 2024 , url=

Reference 81

Resolution
unresolved
no resolver link, observed 2026-08-12T04:51:16.973582Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T04:51:16.973582Z digest=sha256:73de50c5189b7ca36535e3389b89c24457ad7bdbb49df9060f5a1b7303f7cb33

Observation 428c3343-a31c-4b1a-a90f-2fc70ef11344 · outbound

This paper cites Trustworthy LLMs: a Survey and Guideline for Evaluating Large Language Models' Alignment.

From Interpretability to Control: Insights from Six Years of the TrustNLP Workshop Trustworthy LLMs: a Survey and Guideline for Evaluating Large Language Models' Alignment

Reference 82

Resolution
unresolved
no resolver link, observed 2026-08-12T04:51:16.977677Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T04:51:16.977677Z digest=sha256:0264ef31d17dbd80ebc4e68068ff2fdd19f0ea5a3c32a054e80743294c4a663c

Observation 5efb6439-584c-4411-a97b-fb0b4b0a62ce · outbound

This paper cites 2025 , publisher=.

From Interpretability to Control: Insights from Six Years of the TrustNLP Workshop 2025 , publisher=

Reference 83

Resolution
unresolved
no resolver link, observed 2026-08-12T04:51:16.981843Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T04:51:16.981843Z digest=sha256:031bb7cdd9078f1789abe3a3478926628e417ec0a54ee8dda77ed0143e440eb9

Observation 498ceeaf-2846-448f-bc20-7fc603cd67ae · outbound

This paper cites an unresolved cited work.

From Interpretability to Control: Insights from Six Years of the TrustNLP Workshop Unresolved cited work

Reference 84

Resolution
unresolved
no resolver link, observed 2026-08-12T04:51:16.986141Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T04:51:16.986141Z digest=sha256:dc09270b890bf56c563564061871001837dae0c16963af909915b8b31537e171

Observation 37a8edc6-c58f-404c-b0a7-82c03f5dcfb3 · outbound

This paper cites Interpretability Rules: Jointly Bootstrapping a Neural Relation Extractor with an Explanation Decoder.

From Interpretability to Control: Insights from Six Years of the TrustNLP Workshop Interpretability Rules: Jointly Bootstrapping a Neural Relation Extractor with an Explanation Decoder

Reference 85

Resolution
verified exact
doi, observed 2026-08-12T04:51:17.579945Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-12T04:51:16.990589Z digest=sha256:dde85c43d27571ba75a720f04f73feafd6c72aaf1b3375ead156a9955bce35c1

Observation fa99828f-2299-4b74-8df4-932c0a7707fd · outbound

This paper cites Measuring Biases of Word Embeddings: What Similarity Measures and Descriptive Statistics to Use?.

From Interpretability to Control: Insights from Six Years of the TrustNLP Workshop Measuring Biases of Word Embeddings: What Similarity Measures and Descriptive Statistics to Use?

Reference 86

Resolution
verified exact
doi, observed 2026-08-12T04:51:17.897013Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-12T04:51:16.994632Z digest=sha256:5c71052a4240bc5a88baa917e2afe4befd857fd5fb1b3594fc2a8783c3c63b3c

Observation 0d500bcb-aa3f-47bf-870e-b15566842501 · outbound

This paper cites and Kiritchenko, Svetlana and Balkir, Esma.

From Interpretability to Control: Insights from Six Years of the TrustNLP Workshop and Kiritchenko, Svetlana and Balkir, Esma

Reference 87

Resolution
unresolved
no resolver link, observed 2026-08-12T04:51:16.999053Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T04:51:16.999053Z digest=sha256:f1af7e37ecffc668b301ba7913a1433bafca3a34379099dcedac7bc77bd66c49

Observation 977ce2f1-f298-4342-b774-2ceae3df2885 · outbound

This paper cites GPT s Don ' t Keep Secrets: Searching for Backdoor Watermark Triggers in Autoregressive Language Models.

From Interpretability to Control: Insights from Six Years of the TrustNLP Workshop GPT s Don ' t Keep Secrets: Searching for Backdoor Watermark Triggers in Autoregressive Language Models

Reference 88

Resolution
unresolved
no resolver link, observed 2026-08-12T04:51:17.003186Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T04:51:17.003186Z digest=sha256:680ffeef22663fbae1265509df67bb624b5cf400b19f507a18e52c4a2ded1aa1

Observation 855102a8-6674-4972-b483-eab017c5bbfb · outbound

This paper cites Reliability Check: An Analysis of GPT -3's Response to Sensitive Topics and Prompt Wording.

From Interpretability to Control: Insights from Six Years of the TrustNLP Workshop Reliability Check: An Analysis of GPT -3's Response to Sensitive Topics and Prompt Wording

Reference 89

Resolution
unresolved
no resolver link, observed 2026-08-12T04:51:17.007660Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T04:51:17.007660Z digest=sha256:af2e20b58164ae01640ff09982721480b3cc5ab9b0224c8c30c6a9e5fd000d91

Observation a2e7835e-f913-4a22-b938-6112a387e6b0 · outbound

This paper cites Driving Context into Text-to-Text Privatization.

From Interpretability to Control: Insights from Six Years of the TrustNLP Workshop Driving Context into Text-to-Text Privatization

Reference 90

Resolution
unresolved
no resolver link, observed 2026-08-12T04:51:17.011852Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T04:51:17.011852Z digest=sha256:3d9c9ccaa7f945e0531de26d9cd38fe6f00658e0a4ebaf02cc3603e2edaff1d2

Observation a569e7ad-2d73-44ca-8d2e-99b7c9e94b1e · outbound

This paper cites Expanding Scope: Adapting E nglish Adversarial Attacks to C hinese.

From Interpretability to Control: Insights from Six Years of the TrustNLP Workshop Expanding Scope: Adapting E nglish Adversarial Attacks to C hinese

Reference 91

Resolution
unresolved
no resolver link, observed 2026-08-12T04:51:17.016058Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T04:51:17.016058Z digest=sha256:fdafcbb4817f655da119e9eb645d62a4a09221c69b26872314413425ca8dd4bd

Observation 0b3994e2-3d6e-410b-8a35-77b802cde3ba · outbound

This paper cites Flatness-Aware Gradient Descent for Safe Conversational AI.

From Interpretability to Control: Insights from Six Years of the TrustNLP Workshop Flatness-Aware Gradient Descent for Safe Conversational AI

Reference 92

Resolution
unresolved
no resolver link, observed 2026-08-12T04:51:17.020464Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T04:51:17.020464Z digest=sha256:a8f19670044e41ef245f5ee5968b302c4cf8a4fcb2fb5bda464a366c64e83c58

Observation 8a824ce0-ecf5-4508-b1c1-5e2749353908 · outbound

This paper cites PBI -Attack: Prior-Guided Bimodal Interactive Black-Box Jailbreak Attack for Toxicity Maximization.

From Interpretability to Control: Insights from Six Years of the TrustNLP Workshop PBI -Attack: Prior-Guided Bimodal Interactive Black-Box Jailbreak Attack for Toxicity Maximization

Reference 93

Resolution
unresolved
no resolver link, observed 2026-08-12T04:51:17.024785Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T04:51:17.024785Z digest=sha256:bb1d5620bc06cdf86ae382dc62848b5de7400a5c8434f9f8360b1645a152e949

Observation 70156afd-43f2-4b56-9937-21f1a44382fc · outbound

This paper cites Beyond Text-to- SQL for IoT Defense: A Comprehensive Framework for Querying and Classifying IoT Threats.

From Interpretability to Control: Insights from Six Years of the TrustNLP Workshop Beyond Text-to- SQL for IoT Defense: A Comprehensive Framework for Querying and Classifying IoT Threats

Reference 94

Resolution
unresolved
no resolver link, observed 2026-08-12T04:51:17.028942Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T04:51:17.028942Z digest=sha256:d991181eed82a1fa6421797de3a3922f3a1153d26d1bdb7cf740dfa719385b91

Observation 43c6f142-7da7-4d57-95c6-a3c650c01f21 · outbound

This paper cites Minimal Evidence Group Identification for Claim Verification.

From Interpretability to Control: Insights from Six Years of the TrustNLP Workshop Minimal Evidence Group Identification for Claim Verification

Reference 95

Resolution
unresolved
no resolver link, observed 2026-08-12T04:51:17.033333Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T04:51:17.033333Z digest=sha256:5929327c6c2829c9b4c087c53f819e30e64edd2fe152068ca459723d54aea29e

Observation cd32d46d-f950-4093-ad1d-ba9016f142d7 · outbound

This paper cites Estimating Knowledge in Large Language Models Without Generating a Single Token.

From Interpretability to Control: Insights from Six Years of the TrustNLP Workshop Estimating Knowledge in Large Language Models Without Generating a Single Token

Reference 96

Resolution
unresolved
no resolver link, observed 2026-08-12T04:51:17.037344Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T04:51:17.037344Z digest=sha256:8bd653e52229caa3c7e9cea7623b0732b3f7e91652d83f46940b6459793696bf

Observation e0f4c745-e219-4e72-80c3-ed429bbe5883 · outbound

This paper cites Intrinsic Test of Unlearning Using Parametric Knowledge Traces.

From Interpretability to Control: Insights from Six Years of the TrustNLP Workshop Intrinsic Test of Unlearning Using Parametric Knowledge Traces

Reference 97

Resolution
unresolved
no resolver link, observed 2026-08-12T04:51:17.041415Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T04:51:17.041415Z digest=sha256:27002ed803ffa283133e9e3a33f83b4416a6ebda2b40e4dbdf8890439ed2d553

Observation fde013c0-e009-4fed-9000-b955d5bd1970 · outbound

This paper cites Can we trust the evaluation on C hat GPT ?.

From Interpretability to Control: Insights from Six Years of the TrustNLP Workshop Can we trust the evaluation on C hat GPT ?

Reference 98

Resolution
unresolved
no resolver link, observed 2026-08-12T04:51:17.045477Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T04:51:17.045477Z digest=sha256:3986d733b4338e5a1a057da7c84ad48fb4b54fbf2dc01f2d8e6cb5d0aeb7d211

Observation bb79b921-8c81-4537-bcb9-a161c973fd60 · outbound

This paper cites Improving Factuality of Abstractive Summarization via Contrastive Reward Learning.

From Interpretability to Control: Insights from Six Years of the TrustNLP Workshop Improving Factuality of Abstractive Summarization via Contrastive Reward Learning

Reference 99

Resolution
unresolved
no resolver link, observed 2026-08-12T04:51:17.049608Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T04:51:17.049608Z digest=sha256:0101f2342e13d375959f823769c1adb04910017a11bb08956ae1a04f57a2cd6d

Observation 81f29426-d024-4297-bba4-f31f8c9d42e3 · outbound

This paper cites Exploring Causal Mechanisms for Machine Text Detection Methods.

From Interpretability to Control: Insights from Six Years of the TrustNLP Workshop Exploring Causal Mechanisms for Machine Text Detection Methods

Reference 100

Resolution
unresolved
no resolver link, observed 2026-08-12T04:51:17.053733Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T04:51:17.053733Z digest=sha256:1e437ae3f1aaf0446318f7c8ab9ed787f5eb74698c1bca83df3ce29fc2737624

Observation 29cfc788-99b7-49f6-9547-553936eebf07 · outbound

This paper cites On the Robustness of Agentic Function Calling.

From Interpretability to Control: Insights from Six Years of the TrustNLP Workshop On the Robustness of Agentic Function Calling

Reference 101

Resolution
unresolved
no resolver link, observed 2026-08-12T04:51:17.057835Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T04:51:17.057835Z digest=sha256:790dc374a07a7892524c8d4d5aeee622a284a49395215326758d61a53e763247

Observation e27a5b00-a18a-48c3-8fc1-db57c08867f8 · outbound

This paper cites Cross-Task Defense: Instruction-Tuning LLM s for Content Safety.

From Interpretability to Control: Insights from Six Years of the TrustNLP Workshop Cross-Task Defense: Instruction-Tuning LLM s for Content Safety

Reference 102

Resolution
unresolved
no resolver link, observed 2026-08-12T04:51:17.061990Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T04:51:17.061990Z digest=sha256:2e0478fe433037f5dc2eaf5de9080c9372a2425f3d4013f9d0260c0fd6b03210

Observation d9e1eb20-7381-4116-8447-fdb5428c4c1e · outbound

This paper cites Gender Bias in Natural Language Processing Across Human Languages.

From Interpretability to Control: Insights from Six Years of the TrustNLP Workshop Gender Bias in Natural Language Processing Across Human Languages

Reference 103

Resolution
verified exact
doi, observed 2026-08-12T04:51:17.714511Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-12T04:51:17.066494Z digest=sha256:792faae2436ee5ba364244b1d50894d1504bd69c034572c0d672220881aad2d1

Observation c4df4eb7-a254-42fa-a742-7af291b2e8df · outbound

This paper cites Into the Gap between What Language Models Say and What They Know.

From Interpretability to Control: Insights from Six Years of the TrustNLP Workshop Into the Gap between What Language Models Say and What They Know

Reference 104

Resolution
unresolved
no resolver link, observed 2026-08-12T04:51:17.071469Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T04:51:17.071469Z digest=sha256:31a244ea69bedb434b4f87b159c7d6e544cfb8258992b2cdb742b53ffe9137dd

Observation 1dd145f3-52cc-4863-b58e-f57ea78259a9 · outbound

This paper cites The False Sense of Privacy in LLM s: Non-Verbatim Memorization and Semantic Leakage.

From Interpretability to Control: Insights from Six Years of the TrustNLP Workshop The False Sense of Privacy in LLM s: Non-Verbatim Memorization and Semantic Leakage

Reference 105

Resolution
unresolved
no resolver link, observed 2026-08-12T04:51:17.075909Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T04:51:17.075909Z digest=sha256:a2cbe3533f12055e2fb441ddbbe1af630612db17be30f0ef5e27c35d19ce59db

Observation 665d41c0-4741-4d55-bfd6-a3f69d137367 · outbound

This paper cites and Raimundo, Marcos M.

From Interpretability to Control: Insights from Six Years of the TrustNLP Workshop and Raimundo, Marcos M

Reference 106

Resolution
unresolved
no resolver link, observed 2026-08-12T04:51:17.081130Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T04:51:17.081130Z digest=sha256:8e9450e2249d4505010a58a81c0bcf82b21da48ec391568bb302684d3ee00b69

Observation df6bb622-cc75-4d30-a7c2-7f3f0cf08de8 · outbound

This paper cites 2023 , howpublished =.

From Interpretability to Control: Insights from Six Years of the TrustNLP Workshop 2023 , howpublished =

Reference 107

Resolution
unresolved
no resolver link, observed 2026-08-12T04:51:17.085233Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T04:51:17.085233Z digest=sha256:498a8a20debf89af9ad01e4a88f0494735cca4f24cea2d6b379bd8fa8842f22d

Observation b8febde7-8661-42d9-a1ad-9d74840a86c0 · outbound

This paper cites Nature Machine Intelligence , volume=.

From Interpretability to Control: Insights from Six Years of the TrustNLP Workshop Nature Machine Intelligence , volume=

Reference 108

Resolution
unresolved
no resolver link, observed 2026-08-12T04:51:17.089472Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T04:51:17.089472Z digest=sha256:994058ecc3490ebb6412b05bd8aaa8572418ee68052d3597255bd721dd0b3552

Observation 86f3659e-2cca-43b0-90de-d40d432325ad · outbound

This paper cites Formalizing Trust in Artificial Intelligence: Prerequisites, Causes and Goals of Human Trust in AI , year =.

From Interpretability to Control: Insights from Six Years of the TrustNLP Workshop Formalizing Trust in Artificial Intelligence: Prerequisites, Causes and Goals of Human Trust in AI , year =

Reference 109

Resolution
verified exact
arxiv_id_nonexistent, observed 2026-08-12T04:51:18.118140Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-12T04:51:17.093468Z digest=sha256:8afbee57bba7b69b667b4b5037ebb3eae1a64b8c44835a2a80759c531a7b885b

Observation 4ae1e4f8-d7f1-4291-852c-520c7d802d47 · outbound

This paper cites FAccT 2022 , year =.

From Interpretability to Control: Insights from Six Years of the TrustNLP Workshop FAccT 2022 , year =

Reference 110

Resolution
unresolved
no resolver link, observed 2026-08-12T04:51:17.097633Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T04:51:17.097633Z digest=sha256:b3fd3293be2348216f3e5dcd9b4da24b4c1014ad3b098c531185ade1004aaf2f

Observation 0147a468-2552-4eda-a574-a7dee4c5c3ae · outbound

This paper cites SODAPOP : Open-Ended Discovery of Social Biases in Social Commonsense Reasoning Models.

From Interpretability to Control: Insights from Six Years of the TrustNLP Workshop SODAPOP : Open-Ended Discovery of Social Biases in Social Commonsense Reasoning Models

Reference 111

Resolution
verified exact
doi, observed 2026-08-12T04:51:17.453501Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-12T04:51:17.102185Z digest=sha256:c8430b69cd17f5d0af829b3c1f4b065250da796cf012905dfeb3cd3a85928fad

Observation a4a719c5-83ac-4805-bb0e-d9e500235ea1 · outbound

This paper cites F air B elief - Assessing Harmful Beliefs in Language Models.

From Interpretability to Control: Insights from Six Years of the TrustNLP Workshop F air B elief - Assessing Harmful Beliefs in Language Models

Reference 112

Resolution
unresolved
no resolver link, observed 2026-08-12T04:51:17.106463Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T04:51:17.106463Z digest=sha256:599c9bb891518540255d9ea60ba345cd6a03bb1645553c24dc7bd786e35547e6

Observation 6f1b58e1-2ae7-4f1e-b79b-900131422633 · outbound

This paper cites Investigating and Addressing Hallucinations of LLM s in Tasks Involving Negation.

From Interpretability to Control: Insights from Six Years of the TrustNLP Workshop Investigating and Addressing Hallucinations of LLM s in Tasks Involving Negation

Reference 113

Resolution
unresolved
no resolver link, observed 2026-08-12T04:51:17.110769Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T04:51:17.110769Z digest=sha256:e65b7c30d844e62d297bdef7762b245a852d09020de3851a89fe68b346e59e9f

Observation b6c78b3f-8159-465f-af7a-ee3361fa9b89 · outbound

This paper cites Introducing G en C eption for Multimodal LLM Benchmarking: You May Bypass Annotations.

From Interpretability to Control: Insights from Six Years of the TrustNLP Workshop Introducing G en C eption for Multimodal LLM Benchmarking: You May Bypass Annotations

Reference 114

Resolution
unresolved
no resolver link, observed 2026-08-12T04:51:17.114972Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T04:51:17.114972Z digest=sha256:0561850611f71c54a666e5e8f87bce776e136e539393984d410666c6ce6742d3

Observation 1ae8427c-ff08-47d1-a8a4-5c26bd86b8ef · outbound

This paper cites Tell Me Why: Explainable Public Health Fact-Checking with Large Language Models.

From Interpretability to Control: Insights from Six Years of the TrustNLP Workshop Tell Me Why: Explainable Public Health Fact-Checking with Large Language Models

Reference 115

Resolution
unresolved
no resolver link, observed 2026-08-12T04:51:17.119168Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T04:51:17.119168Z digest=sha256:70402b22309061ec50205ef8423c6bc36d8a432075191e2b13b04f211c1efcda

Observation 006a2779-9e8f-4f0a-bcfe-076eb44a3cef · outbound

This paper cites Disentangling Linguistic Features with Dimension-Wise Analysis of Vector Embeddings.

From Interpretability to Control: Insights from Six Years of the TrustNLP Workshop Disentangling Linguistic Features with Dimension-Wise Analysis of Vector Embeddings

Reference 116

Resolution
unresolved
no resolver link, observed 2026-08-12T04:51:17.123464Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T04:51:17.123464Z digest=sha256:8a0496cc76418a2a27d6b6394ceb692d34cf5beadc2a6850c2cd7a4fb67b7c9f

Observation bb1c49fa-76ae-401e-b19e-adb654a3615f · outbound

This paper cites On The Real-world Performance of Machine Translation: Exploring Social Media Post-authors' Perspectives.

From Interpretability to Control: Insights from Six Years of the TrustNLP Workshop On The Real-world Performance of Machine Translation: Exploring Social Media Post-authors' Perspectives

Reference 117

Resolution
unresolved
no resolver link, observed 2026-08-12T04:51:17.127721Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T04:51:17.127721Z digest=sha256:3e1a088678cdf7a68e2db21eece48738b09e6ddffec8ac0daa94d785c421c3fc

Observation 3e020f1c-d738-4c26-8411-4ed746e324a8 · outbound

This paper cites V i B e: A Text-to-Video Benchmark for Evaluating Hallucination in Large Multimodal Models.

From Interpretability to Control: Insights from Six Years of the TrustNLP Workshop V i B e: A Text-to-Video Benchmark for Evaluating Hallucination in Large Multimodal Models

Reference 118

Resolution
unresolved
no resolver link, observed 2026-08-12T04:51:17.131935Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T04:51:17.131935Z digest=sha256:0423f2d22da52260c0041f703189b88e59103dc1ce36f0280652027e050fbaa3

Observation 525ff90f-5571-4d9d-8c82-0760ff712954 · outbound

This paper cites FACTOID : FAC tual en T ailment f O r halluc I nation Detection.

From Interpretability to Control: Insights from Six Years of the TrustNLP Workshop FACTOID : FAC tual en T ailment f O r halluc I nation Detection

Reference 119

Resolution
unresolved
no resolver link, observed 2026-08-12T04:51:17.136528Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T04:51:17.136528Z digest=sha256:fbe684ba4f232a802d24b2e9f8de9afe2e6728ca2cfd9f47bcc86b136b9fbb6f

Observation 90472fb8-a17a-4039-83a5-b2ab7fd06287 · outbound

This paper cites Private Release of Text Embedding Vectors.

From Interpretability to Control: Insights from Six Years of the TrustNLP Workshop Private Release of Text Embedding Vectors

Reference 120

Resolution
verified exact
doi, observed 2026-08-12T04:51:17.384662Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-12T04:51:17.141076Z digest=sha256:6ddb99eccde5111285d5e568d734ffefe7ce835c9d47a6841680317f3c254212

Observation 74282c53-747a-475f-98ee-858b036da7a9 · outbound

This paper cites Challenges in Applying Explainability Methods to Improve the Fairness of NLP Models.

From Interpretability to Control: Insights from Six Years of the TrustNLP Workshop Challenges in Applying Explainability Methods to Improve the Fairness of NLP Models

Reference 121

Resolution
unresolved
no resolver link, observed 2026-08-12T04:51:17.145503Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T04:51:17.145503Z digest=sha256:83096a67bf5d3901c0f91d08f24b5d975327ceb05081213d7cce4cc8fb4ca4ae

Observation 2bb8e51e-980a-4099-8f0a-675e9c86307c · outbound

This paper cites An Encoder Attribution Analysis for Dense Passage Retriever in Open-Domain Question Answering.

From Interpretability to Control: Insights from Six Years of the TrustNLP Workshop An Encoder Attribution Analysis for Dense Passage Retriever in Open-Domain Question Answering

Reference 122

Resolution
unresolved
no resolver link, observed 2026-08-12T04:51:17.149773Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T04:51:17.149773Z digest=sha256:387163bf559b7dbf0f27172d10127f41c1b1fd9f5941043aa767f93ca67c967f

Observation b90ff77c-5bbe-4c31-85ed-7a757a7e202a · outbound

This paper cites A Keyword Based Approach to Understanding the Overpenalization of Marginalized Groups by E nglish Marginal Abuse Models on T witter.

From Interpretability to Control: Insights from Six Years of the TrustNLP Workshop A Keyword Based Approach to Understanding the Overpenalization of Marginalized Groups by E nglish Marginal Abuse Models on T witter

Reference 123

Resolution
unresolved
no resolver link, observed 2026-08-12T04:51:17.154226Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T04:51:17.154226Z digest=sha256:8e14558fa23739c52ae38c5580ef7d5082a47e17c33aacfa5c55370407fe0077

Observation adf32ab9-b7de-43be-b04c-19f465a8f73d · outbound

This paper cites Examining the Causal Impact of First Names on Language Models: The Case of Social Commonsense Reasoning.

From Interpretability to Control: Insights from Six Years of the TrustNLP Workshop Examining the Causal Impact of First Names on Language Models: The Case of Social Commonsense Reasoning

Reference 124

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T04:51:18.541575Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-12T04:51:17.158335Z digest=sha256:31ab70b84a77555eb0047a232774babc8f7a924298b0da51bf2e2bea4b605cd2

Observation 9fa36f13-191e-494a-b98d-78ff65c182fa · outbound

This paper cites An Empirical Study of Metrics to Measure Representational Harms in Pre-Trained Language Models.

From Interpretability to Control: Insights from Six Years of the TrustNLP Workshop An Empirical Study of Metrics to Measure Representational Harms in Pre-Trained Language Models

Reference 125

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T04:51:18.528229Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-12T04:51:17.162510Z digest=sha256:76e2a02a9953b3ea79b41c5a0399981d42ea457a20f554e1a6ffa6faabc1b21a

Observation 341e203f-54ec-4c36-8eb9-9c88008c453a · outbound

This paper cites Beyond T uring: A Comparative Analysis of Approaches for Detecting Machine-Generated Text.

From Interpretability to Control: Insights from Six Years of the TrustNLP Workshop Beyond T uring: A Comparative Analysis of Approaches for Detecting Machine-Generated Text

Reference 126

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T04:51:18.514337Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-12T04:51:17.166956Z digest=sha256:cc13daa72ceedbe203ac523f156735f4d58e2091b2f89977d90a501a85935a6b

Observation 005475b9-56ca-431d-848e-697fccbee092 · outbound

This paper cites Automated Adversarial Discovery for Safety Classifiers.

From Interpretability to Control: Insights from Six Years of the TrustNLP Workshop Automated Adversarial Discovery for Safety Classifiers

Reference 127

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T04:51:18.500532Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-12T04:51:17.171286Z digest=sha256:627f62562b73036fe403f5ee992f2030986a3412303e5969b256d4c89d57d1d8

Observation 7ea20418-09e4-482a-9562-f91ede53ab8f · outbound

This paper cites The Trade-off between Performance, Efficiency, and Fairness in Adapter Modules for Text Classification.

From Interpretability to Control: Insights from Six Years of the TrustNLP Workshop The Trade-off between Performance, Efficiency, and Fairness in Adapter Modules for Text Classification

Reference 128

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T04:51:18.487078Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-12T04:51:17.175925Z digest=sha256:9d0ce808e042dda1ec2f36018068e4821c83af3efd7b58d085882911a5a79c7a

Observation 190fa7b7-a3c4-4d80-876e-5395f7ffa0f7 · outbound

This paper cites On the Interplay between Fairness and Explainability.

From Interpretability to Control: Insights from Six Years of the TrustNLP Workshop On the Interplay between Fairness and Explainability

Reference 129

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T04:51:18.473132Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-12T04:51:17.180480Z digest=sha256:6913ce99e020c0cd4c8d9de8e65afb20a78d898bdd735a1db51219f8f8e8b663

Observation 55ba8d40-09b8-46fd-96da-4a6b551689e9 · outbound

This paper cites F act A lign: Fact-Level Hallucination Detection and Classification Through Knowledge Graph Alignment.

From Interpretability to Control: Insights from Six Years of the TrustNLP Workshop F act A lign: Fact-Level Hallucination Detection and Classification Through Knowledge Graph Alignment

Reference 130

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T04:51:18.458476Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-12T04:51:17.185252Z digest=sha256:c2df5e035ca7c123fb6dfce769a71a941d32f55566e95ddf39b89bc86dcfae9c

Observation 8c75aa3a-5794-4c63-88b4-404d2bc63ee8 · outbound

This paper cites Break the Breakout: Reinventing LM Defense Against Jailbreak Attacks with Self-Refine.

From Interpretability to Control: Insights from Six Years of the TrustNLP Workshop Break the Breakout: Reinventing LM Defense Against Jailbreak Attacks with Self-Refine

Reference 131

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T04:51:18.443147Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-12T04:51:17.190099Z digest=sha256:f73ac85b8ed63efcc44d562f79f8f28e0c86dce28167910fe8e24647c6a9a7e7

Observation 6f66e172-2292-46bb-b6a1-953db8cc11f4 · outbound

This paper cites Ambiguity Detection and Uncertainty Calibration for Question Answering with Large Language Models.

From Interpretability to Control: Insights from Six Years of the TrustNLP Workshop Ambiguity Detection and Uncertainty Calibration for Question Answering with Large Language Models

Reference 132

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T04:51:18.428494Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-12T04:51:17.194565Z digest=sha256:dfc08a29e87d90781c859a792456af4f8df07a0645f1740c40db88abc3d85eb5

Observation 29146040-7016-4cff-8c6c-ea063beffc48 · outbound

This paper cites Error Detection for Multimodal Classification.

From Interpretability to Control: Insights from Six Years of the TrustNLP Workshop Error Detection for Multimodal Classification

Reference 133

Resolution
unresolved
no resolver link, observed 2026-08-12T04:51:17.198931Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T04:51:17.198931Z digest=sha256:ec3d97e3ed50c74c110c79c6de280fb3a00c317b96c1a9bd48cd188740b2ae92

Observation 91c8a12d-937d-4e08-ba69-3eec055fa2a8 · outbound

This paper cites Know What You do Not Know: Verbalized Uncertainty Estimation Robustness on Corrupted Images in Vision-Language Models.

From Interpretability to Control: Insights from Six Years of the TrustNLP Workshop Know What You do Not Know: Verbalized Uncertainty Estimation Robustness on Corrupted Images in Vision-Language Models

Reference 134

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T04:51:18.414831Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-12T04:51:17.203373Z digest=sha256:cb795c28ac57c9f0077103e5c2dfa7463c4de5319478b565631316ab2dde02b5

Observation a7212c6c-36cd-4ffb-b00b-7e7f9fc293b7 · outbound

This paper cites Multi-lingual Multi-turn Automated Red Teaming for LLM s.

From Interpretability to Control: Insights from Six Years of the TrustNLP Workshop Multi-lingual Multi-turn Automated Red Teaming for LLM s

Reference 135

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T04:51:18.400454Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-12T04:51:17.207859Z digest=sha256:3bea8a70abef414a0dc92e2fbb3f950546444a1279b953c1a4df4463b8b45926

Observation 5be57346-eaa7-4f20-8844-ef8f0db1a5bf · outbound

This paper cites Line of Duty: Evaluating LLM Self-Knowledge via Consistency in Feasibility Boundaries.

From Interpretability to Control: Insights from Six Years of the TrustNLP Workshop Line of Duty: Evaluating LLM Self-Knowledge via Consistency in Feasibility Boundaries

Reference 136

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T04:51:18.385972Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-12T04:51:17.212116Z digest=sha256:1fb9de5a1a10c77d6bf332c9d5ba719e6c6ed1a72e47a65d2480cf2bcbcf060e

Observation 4fb8417f-d540-425e-9264-e7743f4588b2 · outbound

This paper cites MoNaCo: More Natural and Complex Questions for Reasoning Across Dozens of Documents.

From Interpretability to Control: Insights from Six Years of the TrustNLP Workshop MoNaCo: More Natural and Complex Questions for Reasoning Across Dozens of Documents

Reference 137

Resolution
unresolved
no resolver link, observed 2026-08-12T04:51:17.216216Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T04:51:17.216216Z digest=sha256:3c5a24c0f588b3884b6f58065f54bffcb39256e3e6722b34751ff0167cf06812

Observation 54f6a97a-3cd1-4e93-afd0-4e4e809588cf · outbound

This paper cites and Aletras, Nikolaos and Ma, Ning.

From Interpretability to Control: Insights from Six Years of the TrustNLP Workshop and Aletras, Nikolaos and Ma, Ning

Reference 138

Resolution
unresolved
no resolver link, observed 2026-08-12T04:51:17.220598Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T04:51:17.220598Z digest=sha256:c601c06aad4f8e52cc61578d1ffeec5491d98c1ef9418f1bbd655461de61e86b

Observation 7046bf1a-afc2-43c3-b9cf-0cafdae14f5b · outbound

This paper cites A Survey on Gender Bias in Natural Language Processing.

From Interpretability to Control: Insights from Six Years of the TrustNLP Workshop A Survey on Gender Bias in Natural Language Processing

Reference 139

Resolution
unresolved
no resolver link, observed 2026-08-12T04:51:17.224833Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T04:51:17.224833Z digest=sha256:927f3f65f6a69aa57154fa69ee92f9422f09a8d9a260cdff4b36a2282bd31180

Observation 443dba07-8f5f-40a6-8d92-ae7ad2edc7fa · outbound

This paper cites Inducing Positive Perspectives with Text Reframing.

From Interpretability to Control: Insights from Six Years of the TrustNLP Workshop Inducing Positive Perspectives with Text Reframing

Reference 140

Resolution
unresolved
no resolver link, observed 2026-08-12T04:51:17.229554Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T04:51:17.229554Z digest=sha256:be8fd58ad6b7efd81fd47ab827e51cb3af3cbf6035527a0741e1d2b99f3ab16f

Observation a71ace5c-74d1-4d18-9a9d-2fe24ea8f402 · outbound

This paper cites The Importance of Modeling Social Factors of Language: Theory and Practice.

From Interpretability to Control: Insights from Six Years of the TrustNLP Workshop The Importance of Modeling Social Factors of Language: Theory and Practice

Reference 141

Resolution
unresolved
no resolver link, observed 2026-08-12T04:51:17.233451Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T04:51:17.233451Z digest=sha256:d91e799cc1ce30ee880e2a4289dca1926dee374b8fd6b90dc32eadffcbedd0a1

Observation 1e9ce44b-5ca1-4dc6-b711-3998d2ef9c6d · outbound

This paper cites Llama 2: Open Foundation and Fine-Tuned Chat Models.

From Interpretability to Control: Insights from Six Years of the TrustNLP Workshop Llama 2: Open Foundation and Fine-Tuned Chat Models

Reference 142

Resolution
unresolved
no resolver link, observed 2026-08-12T04:51:17.237236Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T04:51:17.237236Z digest=sha256:0f4ddb9176adf500fda3cd151a4da982cb6e8bfc52fd1959bdb70d9886a01e95

Observation 258c868f-edfa-44fd-8a50-a06277381181 · outbound

This paper cites 2023 , howpublished =.

From Interpretability to Control: Insights from Six Years of the TrustNLP Workshop 2023 , howpublished =

Reference 143

Resolution
unresolved
no resolver link, observed 2026-08-12T04:51:17.241762Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T04:51:17.241762Z digest=sha256:5d35ad080306953c4a87a5f2b6a6a9e65808bd45906abac006ea0072894645e6

Observation 7650e705-1670-4fde-aa0c-ef1a3c49304a · outbound

This paper cites GPT-4 Technical Report.

From Interpretability to Control: Insights from Six Years of the TrustNLP Workshop GPT-4 Technical Report

Reference 144

Resolution
unresolved
no resolver link, observed 2026-08-12T04:51:17.246222Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T04:51:17.246222Z digest=sha256:8be8244b1985e6db9fcd1848e9a5d9165d43c4c7c659d01b3e0a0d7126b76345

Observation 4a4c4ba2-69bc-47c8-92bb-119dbf2d7bf4 · outbound

This paper cites Strength in Numbers: Estimating Confidence of Large Language Models by Prompt Agreement.

From Interpretability to Control: Insights from Six Years of the TrustNLP Workshop Strength in Numbers: Estimating Confidence of Large Language Models by Prompt Agreement

Reference 145

Resolution
unresolved
no resolver link, observed 2026-08-12T04:51:17.250929Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T04:51:17.250929Z digest=sha256:6fcb003189023410d6f235bed9aecbda0817da563abd75bada6649bd1ef267ca

Observation 4617b8f3-ce41-4840-a759-dd30de6748b9 · outbound

This paper cites On the Intrinsic and Extrinsic Fairness Evaluation Metrics for Contextualized Language Representations.

From Interpretability to Control: Insights from Six Years of the TrustNLP Workshop On the Intrinsic and Extrinsic Fairness Evaluation Metrics for Contextualized Language Representations

Reference 146

Resolution
unresolved
no resolver link, observed 2026-08-12T04:51:17.255122Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T04:51:17.255122Z digest=sha256:828cc40fdacc16f00cf66703dcdeebf106540be1deb80d595f055fd090f44aea

Observation 9b6dd65d-2dd3-4dc9-b107-a6e7c1e3b72c · outbound

This paper cites Pay Attention to the Robustness of C hinese Minority Language Models! Syllable-level Textual Adversarial Attack on T ibetan Script.

From Interpretability to Control: Insights from Six Years of the TrustNLP Workshop Pay Attention to the Robustness of C hinese Minority Language Models! Syllable-level Textual Adversarial Attack on T ibetan Script

Reference 147

Resolution
unresolved
no resolver link, observed 2026-08-12T04:51:17.259117Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T04:51:17.259117Z digest=sha256:62aa1aa9621dddf4eb265f4d1043baa507e3ae02fd472e29b1bc227bb4e3ebf5

Pith citing papers

No inbound Pith citation observations are available.