2026:Team challenges/Team 07E Asie
LexiMap
LexiMap is an analysis tool and Wikimedia gadget that enhances the Wikisource reading experience by streaming linguistic data from Wikidata and Wikimedia lexical resources.
The tool transforms static Wikisource texts into interactive, searchable, and visually enriched documents by automatically analyzing words, identifying their lexical categories (such as nouns, verbs, adjectives, and adverbs), and allowing users to filter and explore text based on linguistic properties.
LexiMap creates a bridge between digital libraries, structured knowledge graphs, and natural language processing, enabling researchers, students, and readers to explore texts in a completely new way.
Problem Statement
Wikisource hosts millions of digitized books and historical documents, but most texts remain static.
Readers currently cannot easily:
- Identify grammatical structures within a text.
- Quickly find all nouns, verbs, or other lexical categories.
- Analyze vocabulary patterns.
- Explore relationships between words and Wikidata concepts.
- Use historical texts for language learning and computational research.
LexiMap solves this by adding an intelligent lexical layer on top of Wikisource content.
Solution Overview
LexiMap will:
- Extract text from Wikisource pages.
- Analyze every word using lexical processing.
- Retrieve structured lexical information from Wikidata/Wikimedia sources.
- Categorize words according to lexical classes.
- Visually map words using color coding.
- Allow users to filter texts by lexical category.
- Provide interactive word-level insights.
Meta Page: https://meta.wikimedia.org/wiki/Leximap
Presentation: https://docs.google.com/presentation/d/18u8mND-9u6IG_gPG15LefpA_FfFWhD3dYp1d27SZYMI/edit?usp=sharing
The Team
Demo Video
Certificate
