A two-stage LLM pipeline generates XPath queries from natural language and sampled web pages, but the reported efficiency gains over the baseline are not backed by any comparative numbers.
Leveraging large language models for web scraping, 2024
1 Pith paper cite this work. Polarity classification is still indexing.
1
Pith paper citing it
citation-role summary
other 1
citation-polarity summary
fields
cs.IR 1years
2024 1verdicts
REJECT 1roles
other 1polarities
unclear 1representative citing papers
citing papers explorer
-
XPath Agent: An Efficient XPath Programming Agent Based on LLM for Web Crawler
A two-stage LLM pipeline generates XPath queries from natural language and sampled web pages, but the reported efficiency gains over the baseline are not backed by any comparative numbers.