Pith. sign in

Paper Citation Record · LEDGER

UserToolBench: A User-Profile-Hidden Benchmark for Personalized Decision Making in Tool-Use LLMs

As of 16 August 2026, this Paper Citation Record lists 35 of 35 outbound references and 0 inbound Pith citation observations for arXiv:2608.10042.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2608.10042 v1

Coverage vector

measured 35 of 35 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-14T04:18:07.485591Z

measured 35 of 35 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-16T06:30:59.297886+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

35 of 35 outbound references displayed

  • verified exact8
  • verified fuzzy13
  • unresolved13
  • parse uncertain0
  • malformed identifier1
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation a0f803dc-53ec-4aa4-86b0-bd2376cd6061 · outbound

This paper cites Tuesday Retail Notes.

UserToolBench: A User-Profile-Hidden Benchmark for Personalized Decision Making in Tool-Use LLMs Tuesday Retail Notes

Reference 1

Resolution
verified exact
raw_fallback, observed 2026-08-14T04:18:07.934741Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-14T04:18:07.440337Z digest=sha256:9773bc1347cafcc8f6c24797dff8993dfa1bc5404d75691c62a80a7c6b37dc1f

Observation db2e51a4-4383-45c8-bf7e-ea8884e0005d · outbound

This paper cites Personalized Language Modeling from Personalized Human Feedback.

UserToolBench: A User-Profile-Hidden Benchmark for Personalized Decision Making in Tool-Use LLMs Personalized Language Modeling from Personalized Human Feedback

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-14T04:18:06.672677Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T04:18:06.672677Z digest=sha256:dc39b9be0644df01762b95c525d62434147a15908d94668863c04d287d5830e8

Observation 517e448d-53ec-4278-a7ac-fb773b0f178b · outbound

This paper cites Aligning LLMs by Predicting Preferences from User Writing Samples.

UserToolBench: A User-Profile-Hidden Benchmark for Personalized Decision Making in Tool-Use LLMs Aligning LLMs by Predicting Preferences from User Writing Samples

Reference 4

Resolution
verified exact
local_arxiv, observed 2026-08-14T04:18:08.623168Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-14T04:18:06.724740Z digest=sha256:5bd349866ce6fe642d61d1c510d3342bfebf1737d2c1d6575c36cf230b41e4f9

Observation a7b0ef19-e581-4183-987f-e346d5b353c1 · outbound

This paper cites API-bank: A comprehensive benchmark for tool-augmented LLMs.

UserToolBench: A User-Profile-Hidden Benchmark for Personalized Decision Making in Tool-Use LLMs API-bank: A comprehensive benchmark for tool-augmented LLMs

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T04:18:09.874925Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-14T04:18:06.774765Z digest=sha256:a17010afcb3ba1eb32adc6a7ccc01901870f0b0d798c30e2baf9bf1549d97eda

Observation ad1d714d-76d2-469d-b3a6-25bc860e3ea8 · outbound

This paper cites Benchmarking LLM Tool-Use in the Wild.

UserToolBench: A User-Profile-Hidden Benchmark for Personalized Decision Making in Tool-Use LLMs Benchmarking LLM Tool-Use in the Wild

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-14T04:18:06.864763Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T04:18:06.864763Z digest=sha256:3b536ca92eb598cbcdaf38c5569557c3e875e1a440127dd5796835f343f67350

Observation 1dd23d7c-8bac-4165-8e68-e4205f2b0193 · outbound

This paper cites $\tau$-bench: A Benchmark for Tool-Agent-User Interaction in Real-World Domains.

UserToolBench: A User-Profile-Hidden Benchmark for Personalized Decision Making in Tool-Use LLMs $\tau$-bench: A Benchmark for Tool-Agent-User Interaction in Real-World Domains

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-14T04:18:06.914755Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T04:18:06.914755Z digest=sha256:757c334c06c4cc76cfb9ff888c2469bffbf2971fd356e12d79520e64fa0f840f

Observation c6b40833-3471-441e-8ff9-c6fa643a70ea · outbound

This paper cites Advancing and Benchmarking Personalized Tool Invocation for LLMs.

UserToolBench: A User-Profile-Hidden Benchmark for Personalized Decision Making in Tool-Use LLMs Advancing and Benchmarking Personalized Tool Invocation for LLMs

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-14T04:18:06.995679Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T04:18:06.995679Z digest=sha256:20c2622461721ca0992b9e9ba108b05af776a8324a9e6d4f869a64def5c88369

Observation a72393e1-6011-4523-b149-226db075730a · outbound

This paper cites Tool- spectrum: Towards personalized tool utilization for large language models.

UserToolBench: A User-Profile-Hidden Benchmark for Personalized Decision Making in Tool-Use LLMs Tool- spectrum: Towards personalized tool utilization for large language models

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T04:18:09.655761Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-14T04:18:07.054748Z digest=sha256:5cafe07aaca8ccbaef84bb253abbb1ed48c344efc44b9241e79d94bf6385f2b4

Observation c3dbc6e2-2c05-4cd2-b0fd-6ffb7a7af9b1 · outbound

This paper cites doi:10.18653/v1/2026.acl-long.370.

UserToolBench: A User-Profile-Hidden Benchmark for Personalized Decision Making in Tool-Use LLMs doi:10.18653/v1/2026.acl-long.370

Reference 12

Resolution
verified exact
doi, observed 2026-08-14T04:18:07.767881Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-14T04:18:07.084830Z digest=sha256:3b4e543bd8308207854a8856d727b20c360ac0a108c30f5d44823d84f354923c

Observation 7ee59956-cf73-4f46-b6a2-0b4c852d1d37 · outbound

This paper cites Fingertip 20k: A benchmark for proactive and personalized mobile llm agents.arXiv preprint arXiv:2507.21071,.

UserToolBench: A User-Profile-Hidden Benchmark for Personalized Decision Making in Tool-Use LLMs Fingertip 20k: A benchmark for proactive and personalized mobile llm agents.arXiv preprint arXiv:2507.21071,

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-14T04:18:07.135485Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T04:18:07.135485Z digest=sha256:e736c9f3d3d725c6e7bfb16795ee83c787bdbdeca8bba7459a940857606796d0

Observation 979fbe54-b07c-42de-b2c6-13e77d830895 · outbound

This paper cites Persona2Web: Benchmarking Personalized Web Agents for Contextual Reasoning with User History.

UserToolBench: A User-Profile-Hidden Benchmark for Personalized Decision Making in Tool-Use LLMs Persona2Web: Benchmarking Personalized Web Agents for Contextual Reasoning with User History

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-14T04:18:07.145058Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T04:18:07.145058Z digest=sha256:596484068ffd68c48899f292c45374c979784d7010aa9e3bba8ba80b45aa4341

Observation 3d496b7a-8c64-40ac-b6e0-1805b5055a90 · outbound

This paper cites Me-agent: A personalized mobile agent with two-level user habit learning for enhanced interaction.arXiv preprint arXiv:2601.20162, 2026a.

UserToolBench: A User-Profile-Hidden Benchmark for Personalized Decision Making in Tool-Use LLMs Me-agent: A personalized mobile agent with two-level user habit learning for enhanced interaction.arXiv preprint arXiv:2601.20162, 2026a

Reference 15

Resolution
verified exact
raw_fallback, observed 2026-08-14T04:18:08.401262Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-14T04:18:07.150631Z digest=sha256:05be9e0d9fda7249542ce2ec517b433e658560a54706524a1a58a4e3601ca755

Observation f8c2895f-6e01-46f0-a2c0-07fe6acf8e51 · outbound

This paper cites ValuePilot: A Two-Phase Framework for Value-Driven Decision-Making.

UserToolBench: A User-Profile-Hidden Benchmark for Personalized Decision Making in Tool-Use LLMs ValuePilot: A Two-Phase Framework for Value-Driven Decision-Making

Reference 16

Resolution
verified exact
local_arxiv, observed 2026-08-14T04:18:08.286973Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-14T04:18:07.165600Z digest=sha256:3ddfd4b8a7f3c0fccc7b6ce163dff96a4ad154903d90612b27fc653ddb727e53

Observation 68011b94-2c97-44dd-8284-0a887c959a75 · outbound

This paper cites Shopsimulator: Evaluating and exploring rl-driven llm agent for shopping assistants.arXiv preprint arXiv:2601.18225, 2026b.

UserToolBench: A User-Profile-Hidden Benchmark for Personalized Decision Making in Tool-Use LLMs Shopsimulator: Evaluating and exploring rl-driven llm agent for shopping assistants.arXiv preprint arXiv:2601.18225, 2026b

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-14T04:18:07.173493Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T04:18:07.173493Z digest=sha256:bac329106df34cce17083be49d546b97e3411ca55383211c1d47c908527464ec

Observation 6932b77f-8bd6-4fed-9041-7eb330819605 · outbound

This paper cites PersonaLLM: Investigating the abil- ity of large language models to express personality traits.

UserToolBench: A User-Profile-Hidden Benchmark for Personalized Decision Making in Tool-Use LLMs PersonaLLM: Investigating the abil- ity of large language models to express personality traits

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T04:18:09.587460Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-14T04:18:07.186039Z digest=sha256:e8c2ed02e68e7c90d8ab7003ce085c7b4ebe4c71bd32f29304bca7588b136bc6

Observation 5f065543-cae1-4658-a75d-66bf879e6adf · outbound

This paper cites URLhttps://aclanthology.org/2024.findings-naacl.229/.

UserToolBench: A User-Profile-Hidden Benchmark for Personalized Decision Making in Tool-Use LLMs URLhttps://aclanthology.org/2024.findings-naacl.229/

Reference 19

Resolution
malformed identifier
no resolver link, observed 2026-08-14T04:18:07.214750Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T04:18:07.214750Z digest=sha256:6f3f6168bdc8a92a4a7a983f0eac16a3162c93fd0fa9e2566737a6ced86c1811

Observation 368143bb-c301-43a6-96f4-1519c277ccdc · outbound

This paper cites an unresolved cited work.

UserToolBench: A User-Profile-Hidden Benchmark for Personalized Decision Making in Tool-Use LLMs Unresolved cited work

Reference 20

Resolution
unresolved
raw_fallback, observed 2026-08-14T04:18:09.469404Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-14T04:18:07.223364Z digest=sha256:80a6037e77c5fd1c98e866bd29a43ae83a67d7086f9fc6d9c92b34ce21fb34bd

Observation 8aa8e873-aeb4-409c-8cd7-21377d555d8b · outbound

This paper cites Qwen Team.

UserToolBench: A User-Profile-Hidden Benchmark for Personalized Decision Making in Tool-Use LLMs Qwen Team

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T04:18:09.406526Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-14T04:18:07.231328Z digest=sha256:a7614b5050aaac05d413fc3862f730061f5c96b00bca9133a16c08407978fd1f

Observation 3d6bb34a-2c46-4c8a-8d39-3e5595bf4294 · outbound

This paper cites DeepSeek-AI.

UserToolBench: A User-Profile-Hidden Benchmark for Personalized Decision Making in Tool-Use LLMs DeepSeek-AI

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T04:18:09.341136Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-14T04:18:07.249330Z digest=sha256:0b73911171d0e906d058f4c5ba6779d5398f72c292ab0a566c0670cfd0c3ceb7

Observation f026ab4a-d16c-4f84-9b91-b675ac11f5f2 · outbound

This paper cites Google DeepMind.

UserToolBench: A User-Profile-Hidden Benchmark for Personalized Decision Making in Tool-Use LLMs Google DeepMind

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T04:18:09.236384Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-14T04:18:07.255437Z digest=sha256:ea4efe31ce19cb4260accb1692edbda49bbd9b5b7b5e0ffaff1dee00500d8886

Observation a8f7a612-f0fd-4e71-a953-bb34be7ba50a · outbound

This paper cites an unresolved cited work.

UserToolBench: A User-Profile-Hidden Benchmark for Personalized Decision Making in Tool-Use LLMs Unresolved cited work

Reference 24

Resolution
unresolved
raw_fallback, observed 2026-08-14T04:18:09.104749Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-14T04:18:07.261202Z digest=sha256:b47073b49e40c31545a9182d9e923461d38da3ab364455aad11c891b88299eed

Observation f26e52b9-39c8-4f0e-8d05-9a286aaab7e4 · outbound

This paper cites Qiqiang Lin, Muning Wen, Qiuying Peng, Guanyu Nie, Junwei Liao, Jun Wang, Xiaoyun Mo, Jiamu Zhou, Cheng Cheng, Yin Zhao, et al.

UserToolBench: A User-Profile-Hidden Benchmark for Personalized Decision Making in Tool-Use LLMs Qiqiang Lin, Muning Wen, Qiuying Peng, Guanyu Nie, Junwei Liao, Jun Wang, Xiaoyun Mo, Jiamu Zhou, Cheng Cheng, Yin Zhao, et al

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-14T04:18:07.268543Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T04:18:07.268543Z digest=sha256:7a87020d609916291db977d523f6dee03425de5bad766ee89e4e1bc3031842ed

Observation 14769fae-8a0c-44f1-8d5c-1eb8751057c8 · outbound

This paper cites Accessed: 2026-05-25.

UserToolBench: A User-Profile-Hidden Benchmark for Personalized Decision Making in Tool-Use LLMs Accessed: 2026-05-25

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T04:18:09.026654Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-14T04:18:07.287222Z digest=sha256:3ed7c744e9aef9baaa57b524efbb52a38d89260d9416357013401f1ab6c7844f

Observation c47f70df-02f7-498f-823e-b26079cea3a4 · outbound

This paper cites co/Team-ACE/ToolACE-2.5-Llama-3.1-8B.

UserToolBench: A User-Profile-Hidden Benchmark for Personalized Decision Making in Tool-Use LLMs co/Team-ACE/ToolACE-2.5-Llama-3.1-8B

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T04:18:08.944279Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-14T04:18:07.296085Z digest=sha256:a8876db63b4f0ca30a86dc4118df98d3bce61d90e9ff3246fd31d06b0ee6ddb6

Observation 105f1510-6014-43e5-a1a1-c0022f71cd55 · outbound

This paper cites Accessed: 2026-05-25.

UserToolBench: A User-Profile-Hidden Benchmark for Personalized Decision Making in Tool-Use LLMs Accessed: 2026-05-25

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T04:18:08.869221Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-14T04:18:07.319485Z digest=sha256:43c5670a071e4bfd73d719d2a885f184b05702cf1653b3d905038c224314e56a

Observation 99a4a60f-b52c-4512-8e2c-3f3a6665f622 · outbound

This paper cites Direct Multi-Turn Preference Optimization for Language Agents.

UserToolBench: A User-Profile-Hidden Benchmark for Personalized Decision Making in Tool-Use LLMs Direct Multi-Turn Preference Optimization for Language Agents

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-14T04:18:07.327158Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T04:18:07.327158Z digest=sha256:d6949a8cc7d75deb5ede6149953f50211d3193048415f4acb27385f5b8967eb5

Observation 06c927ac-55a2-43a5-84b3-8487cd956336 · outbound

This paper cites URL https://aclanthology.org/2026.

UserToolBench: A User-Profile-Hidden Benchmark for Personalized Decision Making in Tool-Use LLMs URL https://aclanthology.org/2026

Reference 30

Resolution
verified exact
doi, observed 2026-08-14T04:18:07.720918Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-14T04:18:07.359764Z digest=sha256:a05ae605aba493373e565a1133893a8e19e936b58b32f83551a1e8f3fbe538be

Observation 32b744bc-020d-46c1-8c25-765f86d01d3a · outbound

This paper cites URL https://aclanthology.org/2026.findings-acl.1080/.

UserToolBench: A User-Profile-Hidden Benchmark for Personalized Decision Making in Tool-Use LLMs URL https://aclanthology.org/2026.findings-acl.1080/

Reference 31

Resolution
verified exact
doi, observed 2026-08-14T04:18:07.689459Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-14T04:18:07.395956Z digest=sha256:394bfbf45e7418a2497b0470a2f9cbf8683215699bf0c5f9d4a6ea1d01e18d9e

Observation 51e3e151-5882-4d0f-bbe5-9ab7da35cd4c · outbound

This paper cites 36Kr GreenLeaf retail rising star.

UserToolBench: A User-Profile-Hidden Benchmark for Personalized Decision Making in Tool-Use LLMs 36Kr GreenLeaf retail rising star

Reference 32

Resolution
verified exact
doi, observed 2026-08-14T04:18:07.595153Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-14T04:18:07.427066Z digest=sha256:8851aeae66b103cb07b2572c829c61e05ef8c863de13ca67fd2adbadd7b1cfdd

Observation 72c0a1b2-46ad-4482-844b-ef89a96ff2d5 · outbound

This paper cites B.3 Quantitative Persona Diversity Audit We complement the field-coverage analysis with a quantitative audit of the ten selected profiles.

UserToolBench: A User-Profile-Hidden Benchmark for Personalized Decision Making in Tool-Use LLMs B.3 Quantitative Persona Diversity Audit We complement the field-coverage analysis with a quantitative audit of the ten selected profiles

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T04:18:08.832172Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-14T04:18:07.458135Z digest=sha256:7dc37b8f06b8ad0b09385f51ef9a6a3abc8112bf4d9ed4e539615e05778b46a8

Observation 07fe565e-04c8-4585-a2c9-2d88c98ca057 · outbound

This paper cites Left: token Jaccard similarity.

UserToolBench: A User-Profile-Hidden Benchmark for Personalized Decision Making in Tool-Use LLMs Left: token Jaccard similarity

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T04:18:08.777001Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-14T04:18:07.485591Z digest=sha256:d566160dfd93aa611cf6e8514044210eac28f236e93832dff58f4fd95a3af48c

Observation 0f9fa615-9808-413b-88f4-21cf915a913d · outbound

This paper cites doi:10.18653/v1/2023.emnlp-main.187.

UserToolBench: A User-Profile-Hidden Benchmark for Personalized Decision Making in Tool-Use LLMs doi:10.18653/v1/2023.emnlp-main.187

Reference 2023

Resolution
unresolved
no resolver link, observed 2026-08-14T04:18:06.818994Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T04:18:06.818994Z digest=sha256:76fb5a214a5b4712c60be00050bea8394733d027d97e9740a4b1ef695ddc2f6e

Observation 2f9c32b5-d519-4eb7-9de1-e782eedbb5e4 · outbound

This paper cites Know me, respond to me: Benchmarking llms for dynamic user profiling and personalized responses at scale.arXiv preprint arXiv:2504.14225,.

UserToolBench: A User-Profile-Hidden Benchmark for Personalized Decision Making in Tool-Use LLMs Know me, respond to me: Benchmarking llms for dynamic user profiling and personalized responses at scale.arXiv preprint arXiv:2504.14225,

Reference 2024

Resolution
unresolved
no resolver link, observed 2026-08-14T04:18:06.605781Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T04:18:06.605781Z digest=sha256:f2cce70ec204f954985bc1225affebb1e014cac9de7a416156208ce1d0931ca8

Observation 27ae0a1f-ea3c-470c-9f0b-b490c0c9ee14 · outbound

This paper cites Personalens: A benchmark for personalization evaluation in conversational ai assistants.

UserToolBench: A User-Profile-Hidden Benchmark for Personalized Decision Making in Tool-Use LLMs Personalens: A benchmark for personalization evaluation in conversational ai assistants

Reference 2025

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T04:18:10.014822Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-14T04:18:06.644752Z digest=sha256:b5198cb3b8786ac3f11d8f012eed99a56b29c54f08570ba2e128aeeaa9457331

Observation 1ef11c0d-b1c0-41ea-bc04-2227aea61bed · outbound

This paper cites Petoolllm: Towards personalized tool learning in large language models.

UserToolBench: A User-Profile-Hidden Benchmark for Personalized Decision Making in Tool-Use LLMs Petoolllm: Towards personalized tool learning in large language models

Reference 2026

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T04:18:09.802674Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-14T04:18:06.948152Z digest=sha256:1357dc3a2ed7f86317b89164e7914995664ed6dfd2663f188d4f75c5bce869bd

Pith citing papers

No inbound Pith citation observations are available.