Pith. sign in

REVIEW 7 cited by

Meltemi: The first open Large Language Model for Greek

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2407.20743 v1 pith:HE43VQPA submitted 2024-07-30 cs.CL

classification cs.CL
keywords meltemigreekcorpusinstructlanguagemodelbeenbillion
verification ladder T0 review T1 audit T2 compute T3 formal

Signed reviews

No signed human review yet.

0 comments
read the original abstract

We describe the development and capabilities of Meltemi 7B, the first open Large Language Model for the Greek language. Meltemi 7B has 7 billion parameters and is trained on a 40 billion token Greek corpus. For the development of Meltemi 7B, we adapt Mistral, by continuous pretraining on the Greek Corpus. Meltemi 7B contains up-to-date information up to September 2023. Furthermore, we have translated and curated a Greek instruction corpus, which has been used for the instruction-tuning of a chat model, named Meltemi 7B Instruct. Special care has been given to the alignment and the removal of toxic content for the Meltemi 7B Instruct. The developed models are evaluated on a broad set of collected evaluation corpora, and examples of prompts and responses are presented. Both Meltemi 7B and Meltemi 7B Instruct are available at https://huggingface.co/ilsp under the Apache 2.0 license.

Discussion (0). Continue with ORCID to comment.

Forward citations

Cited by 7 Pith papers

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. MORFES: A Benchmark for Productive Inflectional Competence in Modern Greek

    cs.CL 2026-07 conditional novelty 6.0 of 10

    MORFES is the first expert-verified Modern Greek productive-inflection benchmark; Sophea-Genesis-1 leads it at 84% per-item production without losing general capability.

  2. CultureMERT: Continual Pre-Training for Cross-Cultural Music Representation Learning

    cs.SD 2025-06 conditional novelty 5.0 of 10

    CultureMERT adapts the MERT-95M music model to Greek, Turkish and Indian traditions via two-stage continual pre-training, improving non-Western auto-tagging by 4.9% on average while keeping Western performance intact.

  3. Krikri: Advancing Open Large Language Models for Greek

    cs.CL 2025-05 conditional novelty 5.0 of 10

    A Greek-focused continual pretraining and post-training pipeline yields Llama-Krikri-8B, which surpasses prior open models on Greek benchmarks and introduces three new Greek evaluation datasets.

  4. Kuwain 1.5B: An Arabic SLM via Language Injection

    cs.CL 2025-04 conditional novelty 5.0 of 10

    Kuwain 1.5B, built by freezing TinyLlama's original layers and training eight added layers plus a 26K Arabic tokenizer, improves Arabic benchmark average from 36.95 to 44.49 while keeping English near 53.28.

  5. Open or Closed LLM for Lesser-Resourced Languages? Lessons from Greek

    cs.CL 2025-01 conditional novelty 5.0 of 10

    Benchmarking Llama-70b against GPT-4o mini on Greek shows task-specific strengths, but the contamination-probe and legal-clustering claims need stronger baselines.

  6. SnakModel: Lessons Learned from Training an Open Danish Large Language Model

    cs.CL 2024-12 conditional novelty 5.0 of 10

    SnakModel, an open Danish 7B LLM, outperforms other Llama2-7B-based models on the ScandEval Danish benchmark, with analyses of training dynamics and data curation.

  7. GR-NLP-TOOLKIT: An Open-Source NLP Toolkit for Modern Greek

    cs.CL 2024-12 conditional novelty 4.0 of 10

    GR-NLP-TOOLKIT is a pip-installable Greek NLP toolkit reporting improved scores over spaCy and Stanza on five tasks, built by fine-tuning GreekBERT and ByT5.

Pith tools