REVIEW 2 cited by
Mini-Giants: "Small" Language Models and Open Source Win-Win
Not yet reviewed by Pith; the record is open.
This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.
SPECIMEN: schema-true, not a live event
T0 review · schema-true
One-sentence machine reading of the paper's core claim.
pith:XXXXXXXX · record.json · timestamp
Signed reviews
read the original abstract
ChatGPT is phenomenal. However, it is prohibitively expensive to train and refine such giant models. Fortunately, small language models are flourishing and becoming more and more competent. We call them "mini-giants". We argue that open source community like Kaggle and mini-giants will win-win in many ways, technically, ethically and socially. In this article, we present a brief yet rich background, discuss how to attain small language models, present a comparative study of small language models and a brief discussion of evaluation methods, discuss the application scenarios where small language models are most needed in the real world, and conclude with discussion and outlook.
Forward citations
Cited by 2 Pith papers
-
Bespoke Visual Assistance: What and How do Blind and Low-Vision People Create with Agentic Programming?
Five blind and low-vision co-designers, using the agentic programming tool ProgramAT, created 37 custom camera-based assistive tools, revealing motivations, iterative strategies, and challenges like model limits and s...
-
Confident or Seek Stronger: Exploring Uncertainty-Based On-device LLM Routing From Benchmarking to Generalization
Uncertainty-based routing thresholds for small-to-large LLM offloading can be bootstrapped from a calibration set built on other datasets, because confidence distributions depend more on the small model and uncertaint...
Discussion (0). Continue with ORCID to comment.