REVIEW 2 cited by
Revisiting Prompt Engineering via Declarative Crowdsourcing
Not yet reviewed by Pith; the record is open.
This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.
SPECIMEN: schema-true, not a live event
T0 review · schema-true
One-sentence machine reading of the paper's core claim.
pith:XXXXXXXX · record.json · timestamp
Signed reviews
read the original abstract
Large language models (LLMs) are incredibly powerful at comprehending and generating data in the form of text, but are brittle and error-prone. There has been an advent of toolkits and recipes centered around so-called prompt engineering-the process of asking an LLM to do something via a series of prompts. However, for LLM-powered data processing workflows, in particular, optimizing for quality, while keeping cost bounded, is a tedious, manual process. We put forth a vision for declarative prompt engineering. We view LLMs like crowd workers and leverage ideas from the declarative crowdsourcing literature-including leveraging multiple prompting strategies, ensuring internal consistency, and exploring hybrid-LLM-non-LLM approaches-to make prompt engineering a more principled process. Preliminary case studies on sorting, entity resolution, and imputation demonstrate the promise of our approach
Forward citations
Cited by 2 Pith papers
-
From Prompt to Pipeline: Large Language Models for Scientific Workflow Development in Bioinformatics
A qualitative study of ten bioinformatics workflows finds LLMs can generate usable Galaxy and Nextflow pipelines, with Gemini best for Galaxy and DeepSeek-V3 best for Nextflow.
-
Quality Control in Open-Ended Crowdsourcing: A Survey
The paper maps quality control techniques for crowdsourcing tasks with large or infinite answer spaces into a two-tiered framework.
Discussion (0). Continue with ORCID to comment.