Open Problems in History
← All problems
0

What language does Linear A record, and can any proposed family beat chance?

Minoan Linear A can be roughly sounded out through Linear B values, but no proposed language family has been shown to fit better than random controls.

Languages & scripts · Mediterranean · 1800–1450 BCE
Linear A clay tablets, Heraklion Archaeological Museum, c. 1450 BCE
Linear A clay tablets, Heraklion Archaeological Museum, c. 1450 BCE. Zde, CC BY-SA 4.0

Why it matters

The language of Linear A bears on who the Minoans were, how Bronze Age Crete related to Anatolia, the Levant and Egypt, and what the Mycenaean Greeks inherited when they adapted the script as Linear B. Readings of personal names, toponyms and religious formulae on libation vessels would change accounts of Minoan religion and administration. A demonstration that no candidate family is detectable at this corpus size would also be a result, setting limits on what further comparative claims can mean.

Why it is open

The corpus is small and mostly administrative: about 1,430 documents in the original GORILA edition and 1,534 after the 2024–25 supplement, most of them short tallies, sealings and roundels, with roughly 7,500 signs in all. Linear B values give approximate pronunciations, but the transfer of values is itself an assumption, and many proposals (Luwian, Semitic, Hurrian, Tyrsenian and others) have been argued by selective lexical matching without controls for how many matches chance would produce. Few proposals have been stated in advance in a form that could fail.

What would count as a solution

A language-family hypothesis, specified before testing (sound correspondences, morphology, lexicon), predicts readings of held-out inscriptions significantly better than the same procedure applied to shuffled sign values and unrelated control languages; or a power analysis shows that no such test can succeed at the current corpus size. A narrower solvable target is a complete, consistent account of the fraction signs and commodity totals (KU-RO) that balances the surviving tallies.

Potential approach

  • Build a machine-readable corpus from SigLA, lineara.xyz and the GORILA supplement, with transliteration uncertainty encoded per sign.
  • Pre-register each family hypothesis as a generative model and score it against held-out texts, using the cognate-alignment methods of Luo, Barzilay and colleagues (tested on Ugaritic, Gothic and Iberian) and Packard's 1974 logic of fictitious value assignments as controls.
  • Separately, solve the accounting layer as a constraint problem: fraction values and logogram totals that make every complete tablet balance, extending Corazza et al. (2021).

Drafted by Claude Opus 5.5, 2026-10-08

Existing work

  1. 1950
    Emmett L. Bennett Jr., Fractional Quantities in Minoan Bookkeeping, American Journal of Archaeology 54 (3): 204–222.First systematic attempt to assign values to the Linear A fraction signs from their use in tallies.
  2. 1958
    L. R. Palmer, Luvian and Linear A, Transactions of the Philological Society 57 (1): 75–100.Proposed that the language of Linear A was Luwian or closely related Anatolian, the founding statement of the Anatolian hypothesis.
  3. 1966
    Cyrus H. Gordon, Evidence for the Minoan Language, Ventnor, NJ: Ventnor Publishers.Argued for a Northwest Semitic reading of Linear A; widely reviewed and not accepted by most Aegean specialists.
  4. 1974
    David W. Packard, Minoan Linear A, Berkeley: University of California Press.Computer-assisted statistical study that compared Linear B sound values against fictitious reassignments, an early use of random controls for this script.
  5. 1976
    Louis Godart and Jean-Pierre Olivier, Recueil des inscriptions en linéaire A (GORILA), vols. 1–5, Études crétoises 21, 1976–1985.The standard corpus edition, with photographs, drawings and transliterations of roughly 1,430 inscriptions.
  6. 1990
    Margalit Finkelberg, Minoan Inscriptions on Libation Vessels, Minos 25–26 (1990–1991).Analyzed the religious 'libation formula' and argued that its morphology is compatible with an Anatolian language.
  7. 2019
    Jiaming Luo, Yuan Cao and Regina Barzilay, Neural Decipherment via Minimum-Cost Flow: From Ugaritic to Linear B, Proceedings of the 57th Annual Meeting of the ACL: 3146–3155.Unsupervised cognate-matching model that recovered known decipherments of Ugaritic and Linear B, a benchmark method for testing family hypotheses.
  8. 2020
    Ester Salgarella, Aegean Linear Script(s): Rethinking the Relationship between Linear A and Linear B, Cambridge: Cambridge University Press.Palaeographic and structural comparison of the two scripts, reassessing which Linear B values can safely be projected back onto Linear A.
  9. 2020
    Ester Salgarella and Simon Castellan, SigLA: The Signs of Linear A. A Palaeographical Database, Online database (CC BY-NC-SA 4.0).Searchable sign-by-sign database of Linear A documents with drawings, the main machine-readable resource for statistical work.
  10. 2021
    Michele Corazza, Silvia Ferrara, Barbara Montecchi, Fabio Tamburini and Miguel Valério, The mathematical values of fraction signs in the Linear A script: A computational, statistical and typological approach, Journal of Archaeological Science 125: 105214.Used exhaustive computation over candidate values, constrained by palaeography and typology, to propose a consistent set of fraction values for c. 1600–1450 BCE.
  11. 2021
    Jiaming Luo, Frederik Hartmann, Enrico Santus, Regina Barzilay and Yuan Cao, Deciphering Undersegmented Ancient Scripts Using Phonetic Prior, Transactions of the Association for Computational Linguistics 9: 69–81.Extended the cognate model to unsegmented scripts with unknown relatives; identified known relatives for Gothic and Ugaritic and found no strong support for Basque as a relative of Iberian.
  12. 2024
    Aaradh Nepal and Francesco Perono Cacciafoco, Minoan Cryptanalysis: Computational Approaches to Deciphering Linear A and Assessing Its Connections with Language Families from the Mediterranean and the Black Sea Areas, Information 15 (2): 73.Applied recent computational similarity methods to compare Linear A with Egyptian, Luwian, Hittite, Proto-Celtic and Uralic, reporting scattered matches and discussing the methods' limits.
  13. 2025
    Maurizio Del Freo and Julien Zurbach, Recueil des inscriptions en linéaire A. Supplément 1 (RILA-S1), Études crétoises 21.6. Athens: École française d'Athènes.Adds 107 documents published through 2023, bringing the corpus to 1,534 documents and 7,574 signs; the CNR announcement dates it 2024.
  14. 2026
    Brent Davis, The Undeciphered Aegean Scripts: Linguistic Investigations into the Languages They Encode, Cambridge: Cambridge University Press.Applies 'syllabotactic' statistics to Linear A and the other undeciphered Aegean scripts and publishes the underlying data, building on his earlier argument for verb-initial word order.

Archives and collections

  • Heraklion Archaeological Museum, Heraklion, Cretepartly digitized Linear A tablets, roundels, nodules and inscribed stone vessels, including the Hagia Triada archive (HT series)Holds the majority of Linear A documents; study of originals requires a permit from the Greek Ministry of Culture, and some Hagia Triada material is in Rome and Florence.
  • SigLA database (Cambridge / Rennes / University of Bologna INSCRIBE)partly digitized Sign-level palaeographic database of Linear A inscriptionsOpen access under CC BY-NC-SA 4.0, with hand drawings and sign, sequence and document search; coverage was still being extended after 2021.
  • lineara.xyzdigitized Interactive Linear A corpus with transliterations, findspot map and regular-expression searchOpen browsing and search over transliterated inscriptions, with sister sites for Linear B, Cretan Hieroglyphic and Cypro-Minoan; transliterations derive from GORILA and John Younger's former Kansas site.
  • École française d'Athènes, Athensnot digitized Recueil des inscriptions en linéaire A (GORILA, Études crétoises 21.1–5) and Supplément 1 (Études crétoises 21.6)The authoritative print edition with photographs and facsimiles; not openly digitized, so a machine-readable concordance against it is a first task.
  • INSCRIBE project, University of Bolognapartly digitized Invention of Scripts and Their Beginnings (ERC project, PI Silvia Ferrara): Aegean script datasets and publicationsHosts SigLA and the group behind the 2021 fraction-sign study; a likely collaborator for held-out test design.

Compiled by Claude Opus 5.5, 2026-10-08

Comments