Ontoterminology Editor (TEDI)
Builds multilingual ontoterminologies: terminologies whose conceptual system is a formal ontology. Six of the seven published datasets were authored in it, then exported to RDF.
Open science
Everything the laboratory produces that can be shared, is. The ontologies and ontoterminologies, deposited with DOIs; the semantic tools that read them; the source code behind those tools; and the published work that explains the thinking. This page is the index to all of it, and to the terms under which each may be reused.
Tools
A suite for working with structured knowledge in the humanities and social sciences. Each tool has its own documentation; the links on the right open the applications themselves.
Builds multilingual ontoterminologies: terminologies whose conceptual system is a formal ontology. Six of the seven published datasets were authored in it, then exported to RDF.
Reads an ontoterminology as a navigable tree. Switch between the concept hierarchy alone and the hierarchy with its instances, then export a standalone copy to share.
Draws knowledge graphs by hand on a canvas: classes, individuals, typed multilingual attributes and labelled relations. Exports to RDF/XML for use anywhere else in the suite.
Explores RDF graphs in the browser with nothing to install. Takes Turtle, RDF/XML, JSON-LD and N-Triples, inspects metadata and namespaces, and queries the graph with SPARQL.
The Python edition, run on your own machine so your files never leave it. Extracts metadata and graph statistics, and includes a built-in SPARQL endpoint.
Queries the published datasets directly in the browser, with worked example queries for each one and export to CSV, JSON and RDF. Nothing to set up.
Analyses Greek corpora: frequency, collocations and co-occurrence networks, plus LLM-backed entity recognition, tagging, sentiment and topic modelling for Ancient and Modern Greek.
Data
Queryable through the SPARQL endpoint.
Queryable through the SPARQL endpoint.
Queryable through the SPARQL endpoint.
Queryable through the SPARQL endpoint.
Queryable through the SPARQL endpoint.
Queryable through the SPARQL endpoint.
Queryable through the SPARQL endpoint.
Outputs
Roche, C., & Papadopoulou, M. MDDT 2024 Proceedings.
Jia, R., Zhang, Z., Jia, Y., Papadopoulou, M., & Roche, C. IEEE Access. doi:10.1109/ACCESS.2024.3487836
Stamatakis, A., Tsakalides, P., & Tamiolaki, M. Frontiers in Political Science 6:1471002.
Tamiolaki, M. Open Research Europe 4:11. doi:10.21956/openreseurope.18050.r38846
Milio, R., Giannadakis, R., & Lourentzaki, A. Proceedings of MDTT 2025, pp. 1–12. CEUR Workshop Proceedings.
Giannadakis, R. TOTh 2024: Terminology & Ontology: Theories and applications. Presses Universitaires Savoie Mont Blanc.
Pinson, S., & Roche, C. Proceedings of the TOTh 2023 Conference. Presses Universitaires Savoie Mont Blanc.
Roche, C. The RDF/XML vocabulary describing the classes and properties of ontoterminologies, as set out in the Concept Theory of Terminology. Underlies the published datasets and OTV.
Roche, C., Papadopoulou, M., Djambian, C., & Vachon, L. (Eds.). Presses Universitaires Savoie Mont Blanc.
Roche, C., Papadopoulou, M., Giannadakis, R., & Vachon, L. (Eds.). Presses Universitaires Savoie Mont Blanc.
Roche, C., & Papadopoulou, M. In New Directions in Digital Terminology Research. Leiden: Brill.
Koidaki, F., & Chatzikyriakidis, S. In H. Bunt (Ed.), Proceedings of the 22nd Joint ACL-ISO Workshop on Interoperable Semantic Annotation and Representation (ISA-22), pp. 68–76. ELRA.
Roche, C., & Katis, E. In Artificial Intelligence in Literacy Education: Foundations, Practices and Innovations. Routledge.
Roche, C. Athens: Herodotos.
Milio, R., Roche, C., & Papadopoulou, M. Proceedings of TOTh 2026, Chambéry. Related: Legal Bodies in Classical Athens.
Giannadakis, R., Papadopoulou, M., & Liu, H. Proceedings of TOTh 2026, Chambéry. Related: Greek and Chinese Philosophers.
Lourentzaki, A., & Papadopoulou, M. In press. Proceedings of TOTh 2025, Chambéry. Related: Hellenistic Events.
Giannadakis, R. TOTh 2025: Terminology & Ontology. Presses Universitaires Savoie Mont Blanc. Related: ALyrA.
Roche, C., Papadopoulou, M., Giannadakis, R., & Vachon, L. (Eds.). Presses Universitaires Savoie Mont Blanc.
Tiktopoulou, K., & Koidaki, F. [Creating a digital corpus from scratch: the case of 19th-century Ionian prose fiction]. In G. Fragkaki & N. Mathioudakis (Eds.), Ψηφιακές εφαρμογές στη νεοελληνική λογοτεχνία. Athens: Gutenberg.
Code
The lab's software is released under the PolyForm Noncommercial 1.0.0: free to use, modify and share for any purpose that is not primarily commercial, with the licence and notices carried forward.
Curated collection of the open datasets and ontologies, supporting digital humanities, computational philology, Greek NLP and cultural heritage data. This repository holds data rather than software, so it carries the dataset terms, not a software licence — see the dataset pages.
Upload RDF/XML, extract metadata and statistics, visualise properties as an interactive graph, run SPARQL queries. Supports TEDI-built ontoterminologies. Documentation.
Desktop application for text analysis: word frequency, named entity recognition, pattern extraction and automatic language detection.
Networks
TALOS Lab is an ordinary member of OPERAS, the European research infrastructure for open scholarly communication in the social sciences and humanities. Announcement ↗
All seven datasets are deposited in the European Commission's Zenodo community, each with a registered DOI and versioned files.
Every deposit is indexed in OpenAIRE, so the data is discoverable through the European open science graph as well as through this site.
TALOS AI4SSH, grant agreement 101087269. Project record on CORDIS, DOI 10.3030/101087269.
Nineteen partner organisations: standards bodies, research groups, sister projects and universities across Europe, China and the Gulf.
TALOS Lab was founded by publication in the Official Greek Government Gazette (FEK 4617B/07-08-2024), which also approved its operational regulation.
Policy
Work produced by the laboratory is published open access wherever the venue allows it.
Datasets are deposited in the European Commission's EU Open Research Repository on Zenodo, each with a registered DOI and versioned files, so that a citation resolves to a fixed version rather than to a moving target. They are built to be FAIR: findable through the DOI and OpenAIRE, accessible over an open SPARQL endpoint, interoperable through RDF and shared vocabularies, and reusable through documented terms and definitions.
Datasets are currently released under CC BY-NC-ND 4.0.
The tools are released under the PolyForm Noncommercial 1.0.0: free to use, modify and share for any purpose that is not primarily commercial, with the licence and notices carried forward. It is source-available rather than open source, since the noncommercial restriction falls outside the Open Source Definition.
The code is published in the TALOS-AI4SSH GitHub organisation.
The TEDI ontoTerminology Editor is the exception. It is distributed by the University of Crete free of charge, but under a personal academic licence rather than an open source one: non-commercial, non-transferable, and not redistributable. Citing it is a condition of use.
Cite the DOI rather than a page on this site. Each dataset page links to its Zenodo record, where the formal citation and every version are held.
TALOS Laboratory
Start typing to search the content of this page.