Pith. sign in

Paper Citation Record · LEDGER

Bridging HCI and AI Research for the Evaluation of Conversational SE Assistants

As of 9 August 2026, this Paper Citation Record lists 41 of 41 outbound references and 3 inbound Pith citation observations for arXiv:2502.07956.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2502.07956 v1

Coverage vector

measured 41 of 41 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-08T11:20:48.524542Z

measured 44 of 44 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 3 of 3 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-06T22:02:40.051068Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Reference resolution

41 of 41 outbound references displayed

  • verified exact0
  • verified fuzzy29
  • unresolved12
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

1
arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Outbound references

Observation 0acf1e46-57f1-48fa-ac7f-dc4a8cb30967 · outbound

This paper cites Current and Future Bots in Software Development,.

Bridging HCI and AI Research for the Evaluation of Conversational SE Assistants Current and Future Bots in Software Development,

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T11:20:50.450929Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-08T11:20:48.347369Z digest=sha256:73b2395540ab99a338b3944238da6fea9d2a6d595acaad5d5e92dfb754dc1084

Observation fee52c28-d20f-43a0-955c-2aced883580c · outbound

This paper cites Large Language Models for Software Engineering: A Systematic Literature Review,.

Bridging HCI and AI Research for the Evaluation of Conversational SE Assistants Large Language Models for Software Engineering: A Systematic Literature Review,

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T11:20:50.436473Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-08T11:20:48.352132Z digest=sha256:8f7f626c5def91ee10af539e13604a166ff71d2ca8f02fb98694de7c39db0f2f

Observation aa231005-e6d7-4c13-91fa-dca5bee4b302 · outbound

This paper cites Bots for pull requests: the good, the bad, and the promising,.

Bridging HCI and AI Research for the Evaluation of Conversational SE Assistants Bots for pull requests: the good, the bad, and the promising,

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T11:20:50.421692Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-08T11:20:48.356446Z digest=sha256:42ab40c6ae4a9f34296c6dc037e02deb710413a3c1a165b5a214428c5f48f16d

Observation 06a0d397-00de-4f6f-a3b7-a144ee72cf08 · outbound

This paper cites Software Engineering and Foundation Models: Insights from Industry Blogs Using a Jury of Foundation Models,.

Bridging HCI and AI Research for the Evaluation of Conversational SE Assistants Software Engineering and Foundation Models: Insights from Industry Blogs Using a Jury of Foundation Models,

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-08T11:20:48.360779Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T11:20:48.360779Z digest=sha256:c3c09c4576ee3ccd59330958b8c0aea84c16b12c41459663b13bd7283b7bf55f

Observation 5a6016ef-1a8a-4d1c-8d23-06cdd89002f9 · outbound

This paper cites Developer Experiences with a Contextualized AI Coding Assistant: Usability, Expectations, and Outcomes,.

Bridging HCI and AI Research for the Evaluation of Conversational SE Assistants Developer Experiences with a Contextualized AI Coding Assistant: Usability, Expectations, and Outcomes,

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T11:20:50.407237Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-08T11:20:48.365987Z digest=sha256:1f2856f020bcea152f8ff479e0829c6e100c4afbe2632c9d846b0131f3cea757

Observation 9b438722-fa06-49ad-9593-7ad670f44c0a · outbound

This paper cites The Programmer’s Assistant: Conversational Interaction with a Large Language Model for Software Development,.

Bridging HCI and AI Research for the Evaluation of Conversational SE Assistants The Programmer’s Assistant: Conversational Interaction with a Large Language Model for Software Development,

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T11:20:50.392556Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-08T11:20:48.370837Z digest=sha256:45448e413d2b17da0bc1b8a3bb699ea0f119c6c17ca2742b7dce57a242004d2b

Observation e6233862-e741-4844-9f2e-c7b8977e7097 · outbound

This paper cites Program Synthesis with Large Language Models.

Bridging HCI and AI Research for the Evaluation of Conversational SE Assistants Program Synthesis with Large Language Models

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-08T11:20:48.376077Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T11:20:48.376077Z digest=sha256:0fc4f304dfadc6611736f26ea53f0f7634567dbe2dc046b9152b4adc8fc27b97

Observation ffae257b-f1f6-4b2d-8ea5-170b6d87331c · outbound

This paper cites How Far Are We? The Triumphs and Trials of Generative AI in Learning Software Engineering,.

Bridging HCI and AI Research for the Evaluation of Conversational SE Assistants How Far Are We? The Triumphs and Trials of Generative AI in Learning Software Engineering,

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T11:20:50.377226Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-08T11:20:48.380824Z digest=sha256:175c1638a165d32685d4cf6c57082ac342944c301d721875c208a816a5abb405

Observation f6839ab3-cb5c-4224-b2eb-3c8843b58ca2 · outbound

This paper cites How practitioners perceive the relevance of software engineering research,.

Bridging HCI and AI Research for the Evaluation of Conversational SE Assistants How practitioners perceive the relevance of software engineering research,

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T11:20:50.362051Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-08T11:20:48.384931Z digest=sha256:9ddcae2070457a7b0460695148f34cd8ba30f6a52400ad1cf0a3763312ba8c85

Observation 304fdf9d-414a-4593-b45d-81785c9357b5 · outbound

This paper cites Programmers Are Users Too: Human-Centered Methods for Improving Programming Tools,.

Bridging HCI and AI Research for the Evaluation of Conversational SE Assistants Programmers Are Users Too: Human-Centered Methods for Improving Programming Tools,

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T11:20:50.346999Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-08T11:20:48.389014Z digest=sha256:8814536829cf14f836743c5aca9221911bbdb6ce91bd322ad99002d6e45eabaa

Observation f9fbf88f-c446-45ef-9d89-6020ab793cfb · outbound

This paper cites What’s (Not) Working in Programmer User Studies?.

Bridging HCI and AI Research for the Evaluation of Conversational SE Assistants What’s (Not) Working in Programmer User Studies?

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T11:20:50.332050Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-08T11:20:48.393136Z digest=sha256:62ad49c4be780c6033d611ba78bfcd38887900771fd396503a64da9bcc10709c

Observation 036bd7f3-460d-4534-8604-0c5b570425fe · outbound

This paper cites Conversational Agents: Goals, Technologies, Vision and Challenges,.

Bridging HCI and AI Research for the Evaluation of Conversational SE Assistants Conversational Agents: Goals, Technologies, Vision and Challenges,

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T11:20:50.317622Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-08T11:20:48.397462Z digest=sha256:f5c7c703ad74a879674e7bca4acf56254fbf35b8e213a17df6086dd9805b0705

Observation 5e1e36a1-be0c-4c4a-8493-f69e7e4d9dcb · outbound

This paper cites Human-Centered Design Recommen- dations for LLM-as-a-judge,.

Bridging HCI and AI Research for the Evaluation of Conversational SE Assistants Human-Centered Design Recommen- dations for LLM-as-a-judge,

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T11:20:50.302949Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-08T11:20:48.401865Z digest=sha256:b904a578f8ad581f507e89b2d2a078aaacc68e24bd5b83f5c2fc4420ab45cc10

Observation e724376b-53ca-4deb-906d-b474f9ea0d31 · outbound

This paper cites PromptMaker: Prompt-based Prototyping with Large Language Models,.

Bridging HCI and AI Research for the Evaluation of Conversational SE Assistants PromptMaker: Prompt-based Prototyping with Large Language Models,

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T11:20:50.288052Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-08T11:20:48.406127Z digest=sha256:540d8276a1bc4bf80f03a6594e8287dfc216b66b199a4c5b809a8ce1835f4b6a

Observation 076daa87-1662-4e0e-a812-6760369f9178 · outbound

This paper cites How NOT To Evaluate Your Dialogue System: An Empirical Study of Unsupervised Evaluation Metrics for Dialogue Response Generation,.

Bridging HCI and AI Research for the Evaluation of Conversational SE Assistants How NOT To Evaluate Your Dialogue System: An Empirical Study of Unsupervised Evaluation Metrics for Dialogue Response Generation,

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T11:20:50.273578Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-08T11:20:48.410343Z digest=sha256:c16bbb994bb94decdf56dc5b818b1a19290644d9a9987e4d483a55aa96a196a7

Observation 4756293d-6c43-4b11-ae54-dc24d1966a0c · outbound

This paper cites SimUser: Generating Usability Feedback by Simulating Various Users Interacting with Mobile Applications,.

Bridging HCI and AI Research for the Evaluation of Conversational SE Assistants SimUser: Generating Usability Feedback by Simulating Various Users Interacting with Mobile Applications,

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T11:20:50.259491Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-08T11:20:48.414486Z digest=sha256:3856006f747dc8fb8f1047e1c8cabd7581fe9d0f0442573ebb4cd15fbf5bb9bc

Observation 90cc9d31-f225-4de6-a773-2efc056098b8 · outbound

This paper cites Leveraging Large Language Models as Simulated Users for Initial, Low-Cost Evaluations of Designed Conversations,.

Bridging HCI and AI Research for the Evaluation of Conversational SE Assistants Leveraging Large Language Models as Simulated Users for Initial, Low-Cost Evaluations of Designed Conversations,

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T11:20:50.246189Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-08T11:20:48.418573Z digest=sha256:7680bebe4db95ad31e456dbd8c929483c082208336832fdcd12ef0ea9f81536a

Observation 5ea89da4-c6ad-4bf3-91d6-dd37042f3801 · outbound

This paper cites Can AI serve as a substitute for human subjects in software engineering research?.

Bridging HCI and AI Research for the Evaluation of Conversational SE Assistants Can AI serve as a substitute for human subjects in software engineering research?

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T11:20:50.232583Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-08T11:20:48.422636Z digest=sha256:77ff36f7ed4b2f9fb38d2481583c00bbeaa4861d36883478e26edd9280dfb86c

Observation a1fefa01-c276-4267-a430-f27290901894 · outbound

This paper cites Judging LLM-as-a-judge with MT-bench and Chatbot Arena,.

Bridging HCI and AI Research for the Evaluation of Conversational SE Assistants Judging LLM-as-a-judge with MT-bench and Chatbot Arena,

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T11:20:50.218481Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-08T11:20:48.426813Z digest=sha256:583d65b900d2b4a6272733b7bc8540878b7e2c8858c4b68833fa7904962f8a72

Observation cbb663ca-e99b-46b8-b065-aa8e1c247304 · outbound

This paper cites Can LLMs Replace Manual Annotation of Software Engineering Artifacts?.

Bridging HCI and AI Research for the Evaluation of Conversational SE Assistants Can LLMs Replace Manual Annotation of Software Engineering Artifacts?

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-08T11:20:48.430892Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T11:20:48.430892Z digest=sha256:26044c101b4255a70e8f1892284b0db89aac690327a69bf2d25bef346c580371

Observation d626cf25-3863-4ae4-9b10-02c7ad4f7e09 · outbound

This paper cites GenderMag: A Method for Evaluating Software’s Gender Inclusiveness,.

Bridging HCI and AI Research for the Evaluation of Conversational SE Assistants GenderMag: A Method for Evaluating Software’s Gender Inclusiveness,

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T11:20:50.204756Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-08T11:20:48.435767Z digest=sha256:293d030db88d979edbeb4ef969aac7f3475f2c0ff9f46b427bdbe33b7ec31e52

Observation 1c3727d9-b534-4319-865a-c8b58c61f72c · outbound

This paper cites How to debug inclusivity bugs?: a debugging process with information architecture,.

Bridging HCI and AI Research for the Evaluation of Conversational SE Assistants How to debug inclusivity bugs?: a debugging process with information architecture,

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T11:20:50.191003Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-08T11:20:48.440124Z digest=sha256:2c0bfb09248c6b9080bbe474beb5a28fe39fdf1ca22349a945dfaece45d23065

Observation 7a7a07ba-5293-4bad-95e5-fa9967550eee · outbound

This paper cites Guidelines for Human-AI Interaction,.

Bridging HCI and AI Research for the Evaluation of Conversational SE Assistants Guidelines for Human-AI Interaction,

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T11:20:50.176398Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-08T11:20:48.444346Z digest=sha256:c1d56193525b92cf0bbfcb1167c1622c892af500d0299ea9dc215015dcbe1737

Observation 355815da-fd27-4ff1-b95c-6f242ad7d458 · outbound

This paper cites Using an LLM to Help With Code Understanding,.

Bridging HCI and AI Research for the Evaluation of Conversational SE Assistants Using an LLM to Help With Code Understanding,

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T11:20:50.162481Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-08T11:20:48.448490Z digest=sha256:831b42967ef61ae1a1f27c34bb503b2d853215ce403e919e53d69579e529a11e

Observation 8f5cc053-c8c4-48e0-874c-ee0b12c65f42 · outbound

This paper cites What You Need is What You Get: Theory of Mind for an LLM-Based Code Understanding Assistant,.

Bridging HCI and AI Research for the Evaluation of Conversational SE Assistants What You Need is What You Get: Theory of Mind for an LLM-Based Code Understanding Assistant,

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T11:20:50.146901Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-08T11:20:48.452788Z digest=sha256:af38ea0d56f3b03df24479e9e4a0ce27d9c15a3c6e0a18a4a3d657841a2cbf28

Observation 32d914df-89e9-4282-a656-032a9425cd88 · outbound

This paper cites Safeguarding Large Language Models: A Survey.

Bridging HCI and AI Research for the Evaluation of Conversational SE Assistants Safeguarding Large Language Models: A Survey

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-08T11:20:48.457016Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T11:20:48.457016Z digest=sha256:8140bba4a80fa5217a5c7d21144a95abfcad786a3f56351c4a8714f2a9044ca3

Observation d58b4392-2252-439d-b65d-61e5d6403d63 · outbound

This paper cites Large Language Models can Accurately Predict Searcher Preferences,.

Bridging HCI and AI Research for the Evaluation of Conversational SE Assistants Large Language Models can Accurately Predict Searcher Preferences,

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T11:20:50.131691Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-08T11:20:48.461914Z digest=sha256:e0b678ab31aa6e5b5130a7da8af2ee237065b1337374a8cd8fc9494486205ca6

Observation 79849c58-eb89-4e69-afdd-0e9260df1382 · outbound

This paper cites Trends, Challenges and Processes in Conversational Agent Design: Exploring Practitioners’ Views through Semi-Structured Interviews,.

Bridging HCI and AI Research for the Evaluation of Conversational SE Assistants Trends, Challenges and Processes in Conversational Agent Design: Exploring Practitioners’ Views through Semi-Structured Interviews,

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T11:20:50.116328Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-08T11:20:48.466258Z digest=sha256:c25c6fce7ebb60bf67e7a71510210de1a822264a8a1b83386b75db62d41b8b88

Observation 4eae2d72-ceb2-425d-a8d9-6cfc3261f661 · outbound

This paper cites Evaluating Large Language Models in Generating Synthetic HCI Research Data: a Case Study,.

Bridging HCI and AI Research for the Evaluation of Conversational SE Assistants Evaluating Large Language Models in Generating Synthetic HCI Research Data: a Case Study,

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T11:20:50.100192Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-08T11:20:48.471029Z digest=sha256:66191fa24e93146e23b8cf440c2e4b1a1b4f583989fd0cd0fbcf8b7a135eb3c3

Observation 2f68c4ab-37c1-46af-bbf9-a05a756d7c10 · outbound

This paper cites Simulating Social Media Using Large Language Models to Evaluate Alternative News Feed Algorithms.

Bridging HCI and AI Research for the Evaluation of Conversational SE Assistants Simulating Social Media Using Large Language Models to Evaluate Alternative News Feed Algorithms

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-08T11:20:48.475363Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T11:20:48.475363Z digest=sha256:8eb040efcaf1d23b9c71eca3f682e9d05435134a01d8897fd56359334a618307

Observation 131f98ee-ef7b-4f3b-8825-b0f763118434 · outbound

This paper cites TeachTune: Reviewing Pedagogical Agents Against Diverse Student Profiles with Simulated Students.

Bridging HCI and AI Research for the Evaluation of Conversational SE Assistants TeachTune: Reviewing Pedagogical Agents Against Diverse Student Profiles with Simulated Students

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-08T11:20:48.480511Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T11:20:48.480511Z digest=sha256:396bb10a4f741951d8b0cced0c4fcde6235bab5cc6e92369aae55cfdeba5bc47

Observation 11503796-0e0e-47c0-8993-61aca8fac571 · outbound

This paper cites Can ChatGPT emulate humans in software engineering sur- veys?.

Bridging HCI and AI Research for the Evaluation of Conversational SE Assistants Can ChatGPT emulate humans in software engineering sur- veys?

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T11:20:50.084921Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-08T11:20:48.485128Z digest=sha256:6d910d43992bd193e16b932c2628b02bc67771bfc907996dd74972393b33960a

Observation b5c40558-960e-416e-b402-753a034c8944 · outbound

This paper cites Arithmetic Control of LLMs for Diverse User Preferences: Directional Preference Alignment with Multi-Objective Rewards.

Bridging HCI and AI Research for the Evaluation of Conversational SE Assistants Arithmetic Control of LLMs for Diverse User Preferences: Directional Preference Alignment with Multi-Objective Rewards

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-08T11:20:48.489366Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T11:20:48.489366Z digest=sha256:a9aa692adbf0f2ed6256097e573cf0bba715cd67d12717a3fd254575c36cda57

Observation c011a80c-ae69-4b41-9bd1-0b3a5b30ab5b · outbound

This paper cites Quantifying the Persona Effect in LLM Simulations.

Bridging HCI and AI Research for the Evaluation of Conversational SE Assistants Quantifying the Persona Effect in LLM Simulations

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-08T11:20:48.494024Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T11:20:48.494024Z digest=sha256:5869272aa161b8c0f588c791ed38f8f8211f9a61c7fa5b7ef5635a654138ba02

Observation 646ea53d-a55e-40eb-87fa-83ceb392d036 · outbound

This paper cites Can LLM be a Personalized Judge?.

Bridging HCI and AI Research for the Evaluation of Conversational SE Assistants Can LLM be a Personalized Judge?

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-08T11:20:48.498755Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T11:20:48.498755Z digest=sha256:bf6e49b9f6dce2a21d7736c9163c57b227f2c5d54e5a95f87c656224c89983fc

Observation 3ec0ac57-e9ce-4eb0-9739-5fb923b9ff43 · outbound

This paper cites Characterizing Software Engineering Work with Personas Based on Knowledge Worker Actions,.

Bridging HCI and AI Research for the Evaluation of Conversational SE Assistants Characterizing Software Engineering Work with Personas Based on Knowledge Worker Actions,

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T11:20:50.069367Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-08T11:20:48.503560Z digest=sha256:9596341730a6175702928424f0656945af3c36cf5481187844c9ebd73d2125e4

Observation 0fe8e11b-5643-49d6-b01c-ce3aa1f24775 · outbound

This paper cites Out of One, Many: Using Language Models to Simulate Human Samples,.

Bridging HCI and AI Research for the Evaluation of Conversational SE Assistants Out of One, Many: Using Language Models to Simulate Human Samples,

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T11:20:50.053933Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-08T11:20:48.507941Z digest=sha256:85f8372a316ad736cdec151eac9d13a66fcdb89c8aa313d280e7bf1559490126

Observation e3d0ab46-adf6-446d-b23a-8cfed4d96147 · outbound

This paper cites Systematic Biases in LLM Simulations of Debates.

Bridging HCI and AI Research for the Evaluation of Conversational SE Assistants Systematic Biases in LLM Simulations of Debates

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-08T11:20:48.512020Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T11:20:48.512020Z digest=sha256:6331447f4604f6a07b2545a813a69cc1ae599b2f8f4943b68a0739db96f4e464

Observation a531e1b5-c86b-43d9-90d3-eee467abd4e1 · outbound

This paper cites Large language models that replace human participants can harmfully misportray and flatten identity groups.

Bridging HCI and AI Research for the Evaluation of Conversational SE Assistants Large language models that replace human participants can harmfully misportray and flatten identity groups

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-08T11:20:48.516339Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T11:20:48.516339Z digest=sha256:48bffe714703d3c0736aac06d4629da0bb53977a1a63d18f04c503eb1ca2baf7

Observation 566caccf-1531-4b4f-a0bd-cb83bb80db71 · outbound

This paper cites Supporting Con- textual Conversational Agent-Based Software Development,.

Bridging HCI and AI Research for the Evaluation of Conversational SE Assistants Supporting Con- textual Conversational Agent-Based Software Development,

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T11:20:50.037619Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-08T11:20:48.520480Z digest=sha256:e6149d5a7f1817952fc2c9e9283f65ffa5dfd5d61718d24aff97ed08ba618feb

Observation 2d9d19aa-3a82-4771-8a84-46adbe57fd25 · outbound

This paper cites Retrieval-Augmented Generation for Large Language Models: A Survey.

Bridging HCI and AI Research for the Evaluation of Conversational SE Assistants Retrieval-Augmented Generation for Large Language Models: A Survey

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-08T11:20:48.524542Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T11:20:48.524542Z digest=sha256:dd2ce5a01018cd93a49c7f6bf497202223103250d13964d375a41f3fd42f82cb

Pith citing papers

Observation 2fa50049-4b1e-41f4-801b-d5d2a67165f4 · inbound

DICE-BENCH: Evaluating the Tool-Use Capabilities of Large Language Models in Multi-Round, Multi-Party Dialogues cites this paper.

DICE-BENCH: Evaluating the Tool-Use Capabilities of Large Language Models in Multi-Round, Multi-Party Dialogues Bridging HCI and AI Research for the Evaluation of Conversational SE Assistants

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-06T22:02:40.051068Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T22:02:40.051068Z digest=sha256:f510cc42b1a5c885c0bf49df09ab557a5022f520384190ea824e3689af93767a

Observation a9d9e85b-fd2a-4583-b048-160051c28d7a · inbound

Evaluating the Effectiveness of Large Language Models in Solving Simple Programming Tasks: A User-Centered Study cites this paper.

Evaluating the Effectiveness of Large Language Models in Solving Simple Programming Tasks: A User-Centered Study Bridging HCI and AI Research for the Evaluation of Conversational SE Assistants

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-06T19:59:19.373313Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:59:19.373313Z digest=sha256:ea6f3ccd7b84b82f6caae23fe6f398d1870c31437d2e4144c6748642b5f69f15

Observation 8d3fe331-9068-4595-abd9-058aa98381d2 · inbound

Personalizing LLM-Based Conversational Programming Assistants cites this paper.

Personalizing LLM-Based Conversational Programming Assistants Bridging HCI and AI Research for the Evaluation of Conversational SE Assistants

Reference 19

Resolution
verified exact
arxiv_id, observed 2026-05-10T14:35:32.042898Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-10T14:33:15.633633Z digest=sha256:c1dca29c7ee8e6111e316cf5b79e3563a316a9e7be70423fc6e6c8ac9d21fc3c