REVIEW 1 cited by
Forecasting Credit Ratings: A Case Study where Traditional Methods Outperform Generative LLMs
Not yet reviewed by Pith; the record is open.
This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.
SPECIMEN: schema-true, not a live event
T0 review · schema-true
One-sentence machine reading of the paper's core claim.
pith:XXXXXXXX · record.json · timestamp
read the original abstract
Large Language Models (LLMs) have been shown to perform well for many downstream tasks. Transfer learning can enable LLMs to acquire skills that were not targeted during pre-training. In financial contexts, LLMs can sometimes beat well-established benchmarks. This paper investigates how well LLMs perform in the task of forecasting corporate credit ratings. We show that while LLMs are very good at encoding textual information, traditional methods are still very competitive when it comes to encoding numeric and multimodal data. For our task, current LLMs perform worse than a more traditional XGBoost architecture that combines fundamental and macroeconomic data with high-density text-based embedding features.
Forward citations
Cited by 1 Pith paper
-
When Dimensionality Hurts: The Role of LLM Embedding Compression for Noisy Regression Tasks
Compressing LLM text embeddings with an autoencoder to about 8 dimensions improves stock return prediction, but this benefit disappears on high-signal tasks, and sentiment features seem to work mainly because of compression.
Discussion (0). Continue with ORCID to comment.