REVIEW 21 cited by
NLTK: The Natural Language Toolkit
Not yet reviewed by Pith; the record is open.
This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.
SPECIMEN: schema-true, not a live event
T0 review · schema-true
One-sentence machine reading of the paper's core claim.
pith:XXXXXXXX · record.json · timestamp
read the original abstract
NLTK, the Natural Language Toolkit, is a suite of open source program modules, tutorials and problem sets, providing ready-to-use computational linguistics courseware. NLTK covers symbolic and statistical natural language processing, and is interfaced to annotated corpora. Students augment and replace existing components, learn structured programming by example, and manipulate sophisticated models from the outset.
Forward citations
Cited by 21 Pith papers
-
Recovering Latent Structures after Variational Bayesian Variable Selection: Fit Assessment and Factor-Number Selection in Partially Exploratory Factor Analysis
A scale-free gain rule applied to variational ELBO paths recovers true factor dimensionality in partially exploratory factor analysis where raw information criteria over-factor.
-
LightSTAR: Efficient Visual Document Retrieval via Lightweight Selection with Vision-Adaptive Refinement
LightSTAR achieves state-of-the-art accuracy in visual document retrieval by decomposing the task into LLM-free high-recall candidate selection and vision-adaptive semantic refinement on candidates, cutting end-to-end...
-
A Multi-Agent Framework for Feature-Constrained Difficulty Control in Reading Comprehension Item Generation
MAFIG is a multi-agent framework that uses LLM agents and evaluators to generate reading comprehension items with significantly higher adherence to specified feature constraints than single-agent baselines.
-
Medical Image De-Identification Resources: Synthetic DICOM Data and Tools for Validation
The MIDI resource provides a 53,581-image synthetic DICOM dataset with known PHI/PII insertions and an answer-key-driven validation script for benchmarking de-identification workflows.
-
Mitigating Object Hallucinations via Sentence-Level Early Intervention
SENTINEL reduces MLLM object hallucinations by over 90% via sentence-level early intervention with detector-bootstrapped preference data and C-DPO loss, outperforming prior SOTA on hallucination and capability benchmarks.
-
A Lightweight Method to Disrupt Memorized Sequences in LLM
A decoding-time intervention that substitutes a small model's probabilities for common function words into a large model's output reduces exact training-data recall by up to 10x with minimal measured quality loss.
-
Speculative Prefill: Turbocharging TTFT with Lightweight and Training-Free Token Importance Estimation
A training-free method that selects a subset of prompt tokens with a small speculator model to accelerate LLM prefill, yielding up to 7.66x TTFT speedup on Llama-3.1-405B.
-
Human-Guided Image Generation for Expanding Small-Scale Training Image Datasets
A human-guided image-generation tool with contrastive multi-modal projection and sample-level prompt feedback lifted classification accuracy from 48.45% to 81.80% in a 10-class pet case study.
-
What makes a good metric? Evaluating automatic metrics for text-to-image consistency
None of the four tested text-to-image consistency metrics satisfies all proposed validity criteria, and the VQA-based metrics appear to rely largely on text priors such as yes-bias.
-
Smaug: Fixing Failure Modes of Preference Optimisation with DPO-Positive
DPOP is a new loss function that prevents DPO from lowering preferred response likelihoods and outperforms standard DPO on diverse datasets, MT-Bench, and enables Smaug-72B to exceed 80% on the Open LLM Leaderboard.
-
Fast and Accurate Capitalization and Punctuation for Automatic Speech Recognition Using Transformer and Chunk Merging
Overlapping chunks plus a merging rule improve Transformer-based capitalization and punctuation restoration for speech transcripts, but the claimed edge over prior systems is not directly measured.
-
DocRetriever: A Plug-and-Play Framework for Multimodal Document Retrieval with Comprehensive Benchmark
DocRetriever introduces a framework using layout-aware sparse embeddings for hybrid encoding without OCR and a generalizable reasoning-augmented reranker for few-shot settings, plus the MultiDocR benchmark for evaluation.
-
An Exploration of Internal States in Collaborative Problem Solving
In a Lego-based collaborative task, participants' retrospective verbal reports show distinct linguistic patterns, with positive emotion labels such as 'Engaged' and 'Optimistic' appearing most frequently.
-
Fake News Detection After LLM Laundering: Measurement and Explanation
LLM paraphrasing of fake news degrades detector performance across 17 detectors, with Pegasus evading best and a sentiment shift that BERTScore fails to capture.
-
LongKey: Keyphrase Extraction for Long Documents
A long-document keyphrase extractor using Longformer, convolution-based n-gram embeddings, and max-pooling over occurrences outperforms prior extractors on LDKP and most zero-shot datasets.
-
A Bug or a Suggestion? An Automatic Way to Label Issues
The paper claims an attention-based BiLSTM with k-NN label correction achieves 85.6% F-measure for cross-platform bug/non-bug issue classification.
-
A Weakly-Supervised Attention-based Visualization Tool for Assessing Political Affiliation
A BiLSTM with self-attention, trained on noisy labels from Twitter user descriptions, can produce low-dimensional projections that let a user spot mislabeled political accounts, with MDS reported as the fastest projec...
-
What Differentiates Educational Literature? A Multimodal Fusion Approach of Transformers and Computational Linguistics
A multimodal model (ELECTRA plus a linguistic-feature network) reportedly classifies literature into UK Key Stages with F1 0.996, though the evaluation split may leak book-level information.
-
A Deep Learning Approach for Tweet Classification and Rescue Scheduling for Effective Disaster Management
An attention-based deep learning model with handcrafted features classifies disaster tweets into rescue-need categories, and a priority-aware multi-task scheduler orders rescue missions.
-
Scalable AI-Driven Analytics for User Engagement and Stance Detection on Social Media
A scalable service framework combining standard NLP components is applied to 7M YouTube comments, revealing that conspiracy videos receive up to 70% of engagement in the first week and that most users express favorabl...
-
Performance Evaluation of Supervised Machine Learning Techniques for Efficient Detection of Emotions from Online Content
A benchmark of eight standard classifiers on the ISEAR emotion dataset shows the back-propagation neural network (71.27% accuracy) and logistic regression (66.58%) score highest, but reporting errors in the paper make...
Discussion (0). Continue with ORCID to comment.