Jump to content

2026:Team challenges/Team 07E Asie

From Wikimania
Welcome on the page of Team 07E Asie !
Stream Data with Wikidata

LexiMap

Leximap

LexiMap is an analysis tool and Wikimedia gadget that enhances the Wikisource reading experience by streaming linguistic data from Wikidata and Wikimedia lexical resources.

The tool transforms static Wikisource texts into interactive, searchable, and visually enriched documents by automatically analyzing words, identifying their lexical categories (such as nouns, verbs, adjectives, and adverbs), and allowing users to filter and explore text based on linguistic properties.

LexiMap creates a bridge between digital libraries, structured knowledge graphs, and natural language processing, enabling researchers, students, and readers to explore texts in a completely new way.

Problem Statement

Wikisource hosts millions of digitized books and historical documents, but most texts remain static.

Readers currently cannot easily:

  • Identify grammatical structures within a text.
  • Quickly find all nouns, verbs, or other lexical categories.
  • Analyze vocabulary patterns.
  • Explore relationships between words and Wikidata concepts.
  • Use historical texts for language learning and computational research.

LexiMap solves this by adding an intelligent lexical layer on top of Wikisource content.

Solution Overview

LexiMap will:

  1. Extract text from Wikisource pages.
  2. Analyze every word using lexical processing.
  3. Retrieve structured lexical information from Wikidata/Wikimedia sources.
  4. Categorize words according to lexical classes.
  5. Visually map words using color coding.
  6. Allow users to filter texts by lexical category.
  7. Provide interactive word-level insights.

Meta Page: https://meta.wikimedia.org/wiki/Leximap

Presentation: https://docs.google.com/presentation/d/18u8mND-9u6IG_gPG15LefpA_FfFWhD3dYp1d27SZYMI/edit?usp=sharing

The Team

Demo Video

demo video for the leximap tool/gadget

Certificate