Pith. sign in

REVIEW 4 cited by

pysentimiento: A Python Toolkit for Opinion Mining and Social NLP tasks

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2106.09462 v3 pith:WGTX7CAF submitted 2021-06-17 cs.CL

classification cs.CL
keywords socialtaskspythoncomprehensiveenglishissueslanguageslibrary
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

In recent years, the extraction of opinions and information from user-generated text has attracted a lot of interest, largely due to the unprecedented volume of content in Social Media. However, social researchers face some issues in adopting cutting-edge tools for these tasks, as they are usually behind commercial APIs, unavailable for other languages than English, or very complex to use for non-experts. To address these issues, we present pysentimiento, a comprehensive multilingual Python toolkit designed for opinion mining and other Social NLP tasks. This open-source library brings state-of-the-art models for Spanish, English, Italian, and Portuguese in an easy-to-use Python library, allowing researchers to leverage these techniques. We present a comprehensive assessment of performance for several pre-trained language models across a variety of tasks, languages, and datasets, including an evaluation of fairness in the results.

Discussion (0). Continue with ORCID to comment.

Forward citations

Cited by 4 Pith papers

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. Examining Spanish Counseling with MIDAS: a Motivational Interviewing Dataset in Spanish

    cs.CL 2025-02 conditional novelty 6.0 of 10

    MIDAS, a new expert-annotated Spanish motivational interviewing dataset, reveals language-specific counselor behaviors and supports Spanish-language behavior classification.

  2. A Durability and Cross-Language Transfer Benchmark for a Validated Teaching-Feedback Classification Protocol

    cs.CL 2026-07 accept novelty 5.0 of 10

    The teaching-feedback classification protocol remains durable across three representation generations and transfers to English sentiment, so model choice is a deployment decision.

  3. Specializing General-purpose LLM Embeddings for Implicit Hate Speech Detection across Datasets

    cs.CL 2025-08 conditional novelty 5.0 of 10

    Fine-tuning large general-purpose text embeddings with a simple instruction yields state-of-the-art implicit hate speech detection, with up to 20.35 point cross-dataset F1 gains.

  4. psytechlab at CLPsych 2026: Utilising Natural Language Processing methods and Large Language Models for Social Media Text Analysis

    cs.CL 2026-07 conditional novelty 3.0 of 10

    A standard NLP/LLM pipeline for CLPsych 2026 self-state, change-detection, and timeline-summarization tasks reached top-tier consistency metrics on summaries and mid-level ranks elsewhere.

Pith tools