Ìmọ̀dòtun AI
Ìmọ̀dòtun AI
Preserving Wisdom. Empowering Generations.
Platform Features

AI Tools Built for
Cultural Preservation

Six core AI-powered features — from intelligent OCR and multi-language translation to semantic search and Traditional Knowledge labeling — working together as a complete indigenous knowledge preservation pipeline.

Explore Features Try Live Demo
AI technology
All Features

Six Tools. One Pipeline.

Each feature is designed to work independently and as part of the automated end-to-end archiving pipeline

Digital Archive
A searchable, filterable digital vault for all manuscripts, research papers, and cultural documents. Organised by category, language, community, and date.
Learn More
Multi-Language Translation
AI-powered bidirectional translation across English and 11+ Nigerian languages — with a culturally aware glossary layer for indigenous terminology.
Learn More
Smart OCR
Extracts text from scanned manuscripts, handwritten notes, and image-based documents using Tesseract and/or Google Vision API with high confidence scoring.
Learn More
AI Categorization
Automatically suggests the knowledge category, originating community, and Traditional Knowledge (TK) labels for every uploaded item. Admins confirm or correct.
Learn More
Semantic Search
Goes beyond keyword matching — searches the meaning and context of archived content, returning relevant results even when exact words don't match.
Learn More
Cultural Integrity System
Community consent tracking, TK label management, ethical use guidelines, and Cultural Sensitivity review — built into the core workflow, not bolted on.
Learn More
How Features Connect

The Automated Archiving Pipeline

Every uploaded manuscript flows through this pipeline automatically — only publishing once it has passed human review

📤
Upload
Admin uploads PDF, DOCX, or scanned image with metadata and consent confirmation
👁️
OCR Extract
Text extracted from the file — parsed from native PDFs or recognised via OCR for scans
🌐
Translate
Content translated to all target languages with indigenous glossary post-processing
🏷️
Categorize
AI suggests knowledge category, community, and Traditional Knowledge (TK) labels
🔍
Review
Human reviewer checks accuracy, edits AI output, confirms cultural integrity
🌍
Publish
Approved item goes live in the public archive — searchable by anyone worldwide
Feature 01

Digital Archive

The heart of the platform — a permanent, searchable digital vault for all forms of indigenous knowledge. Every archived item is tagged with language, category, community of origin, Traditional Knowledge labels, and publication date, making the entire corpus discoverable and filterable.

Full-text search across all archived manuscripts and translations
Filter by knowledge category, language, community, date range
Download original document, translated version, or both
Each item shows source language, TK labels, and consent status
Public access — no login required to browse or read
Indexed with F1 score ≥ 0.80 against gold-standard metadata
Archive Search
Ìtàn Ede Kingdom — Oral History Manuscript
Yorùbá → English · Oral History · Ede, Osun State
Igbo Medicinal Herb Compendium Vol. 3
Igbo → English · Traditional Medicine · Anambra
Tiv Agricultural Practices: Seasonal Calendar
Tiv → English · Agriculture · Benue State
Fulfulde Governance & Customary Law Codex
Fulfulde → English · Governance · Adamawa

Feature 02

Multi-Language Translation

AI-powered translation across English and all major Nigerian languages, with a custom post-processing glossary layer that corrects AI outputs for culturally specific terms, place names, and indigenous concepts that generic translation models often mistranslate. The glossary is editable by admins and grows with every review.

Dynamic source/target language selector — any pair, any direction
Cultural glossary layer for 11+ Nigerian languages
BLEU score ≥ 25 / BERT score ≥ 0.85 target benchmarks
Human review step before translated text is published
Stores both AI translation and human-corrected version
Live translation widget available to public on AI Tools page
AI Translator
Yorùbá (Source)
Ìmọ̀ àgbàdo ni ipilẹ̀ igbẹ̀kẹ̀lé àwùjọ. Àwọn ẹ̀kọ́ àgbà jẹ́ ohun ìjókòó nínú ìtàn ìlú.
English (Translated)
Knowledge of corn cultivation is the foundation of community trust. The teachings of the elders are the cornerstone of the town's history.
BLEU: 27.3 ✓ BERT: 0.87 ✓ Glossary: 2 terms

Feature 03

Smart OCR

Many invaluable indigenous knowledge manuscripts exist only as physical documents — handwritten, typed, or photographed. Ìmọ̀dòtun AI's OCR engine extracts text from these materials automatically, detecting whether a file needs OCR or can be parsed directly, and reporting a confidence score so reviewers know where to focus corrections.

PDF text parsing for native digital documents (no OCR needed)
Tesseract OCR engine for scanned images and photographs
Google Vision API option for complex or degraded documents
Confidence score per page — flags low-confidence sections for review
Supports PDF, DOCX, JPG, PNG, TIFF formats
Extracted text stored separately from original file
Manuscript scanning

Feature 04

AI Categorization

After text is extracted, the AI analyses the content and automatically suggests the most appropriate knowledge category, the likely originating community, and a set of Traditional Knowledge (TK) labels. Admins review and approve or correct these before the item is published — maintaining ≥ 90% TK label accuracy as a platform benchmark.

8 knowledge categories — admin-editable taxonomy
Community/ethnic group auto-suggestion with confidence score
Traditional Knowledge (TK) label generation per ENRICH standard
≥ 90% TK label coverage benchmark across the archive
Human reviewer approves/rejects each AI suggestion
Auto-generated keywords used to power semantic search
AI Categorizer — Suggestions
AI Suggested Category
📜 Oral History & Folklore 94% confidence
Community of Origin
Yorùbá — Ede Community, Osun State 88% confidence
Generated TK Labels
TK Oral TK Community TK Attribution TK Verified


Feature 06

Cultural Integrity System

Not a policy document — a built-in workflow. Before any item can be uploaded, the admin must confirm community consent. Before it can be published, a reviewer must confirm TK labels and cultural accuracy. Before it appears publicly, a Super Admin can make a final call. 100% cultural integrity is a benchmark, not an aspiration.

Consent confirmation required at upload — cannot be skipped
TK label review at every publication step
Cultural sensitivity flags in the reviewer workflow
Consent documentation stored with every archive record
Items can be unpublished at any time by request
100% cultural integrity tracking in the analytics dashboard
Upload — Consent Step
File uploaded & validated Done
Metadata form completed Done
Community consent confirmed Required
TK labels reviewed & approved Pending
Published to public archive Pending
Technical Specifications

What's Under the Hood

A transparent view of the technology stack, AI services, and performance targets behind each platform feature

Feature Technology / Service Performance Target Input Formats Human Review
Digital Archive MySQL + PHP PDO + FULLTEXT index F1 ≥ 0.80 All types
Multi-Language Trans. Google Translate API + custom glossary BLEU ≥ 25 / BERT ≥ 0.85 Extracted text
Smart OCR Tesseract / Google Vision API ≥ 85% page confidence PDF, JPG, PNG, TIFF
AI Categorization NLP keyword analysis + taxonomy matching TK coverage ≥ 90% Extracted text
Semantic Search MySQL FULLTEXT + vector similarity (Phase 5) Relevance ≥ top-5 All archive content
Cultural Integrity Built-in workflow gates + consent DB fields 100% coverage All uploads
FAQ

Frequently Asked Questions

Everything you need to know about how the platform works, who can access it, and how content is protected

Ask a Question

Only authenticated admin accounts can upload content. Admin accounts are created by the Super Admin (Ede Polytechnic staff). Public visitors can browse and search the archive but cannot upload.

The platform accepts PDF, DOCX, DOC (native digital documents), and JPG, PNG, TIFF (scanned images and photographs). Maximum file size is 50MB per upload.

Every item goes through a human review stage before publication. The reviewer can edit the AI-generated translation, add corrections, and flag any culturally incorrect interpretations. The corrected version is stored separately and used as the published output.

Yes — the public archive is open to anyone, worldwide, without requiring a login. Visitors can browse, filter, search, and read all published items. Only the admin dashboard and upload functions require authentication.

Consent is captured at the upload stage — the admin must confirm that the originating community has given permission for the material to be archived and published. This confirmation is stored in the database with every record and is visible on each public archive entry.

Yes. Admins and Super Admins can unpublish or delete any item at any time — for example, if a community withdraws consent or if an error is discovered. The platform supports full lifecycle management of archived content.

See These Features
In Action

Try the live AI Translator, browse the public archive, or contact us to learn more about the platform.

Try Live AI Tools Browse the Archive