Skip to content
#

nlp-dataset

Here are 46 public repositories matching this topic...

ancient-greek-texts

CC0 corpus of 1,354 ancient Greek authors and 4,053 works — Homer through late antiquity. Philosophy, history, drama, lyric, medicine, mathematics, rhetoric, and the fragmentary traditions, in clean Unicode Greek. PDF, Markdown, plain text, and JSON. Data store for Eulogikon (https://eulogikon.org). AI training permitted.

  • Updated Sep 3, 2026

Open dataset of Italian case law in Markdown: 22,000+ Constitutional Court decisions (1956 to today), Cassazione Massimario digest, merit-court radar. Verifiable official sources, CC BY 4.0. · Banca dati aperta di giurisprudenza italiana: Consulta, Cassazione e merito, con fonti ufficiali verificabili.

  • Updated Sep 7, 2026
  • Python

Add this topic to your repo

To associate your repository with the nlp-dataset topic, visit your repo's landing page and select "manage topics."

Learn more