ISC Digital Humanities
  • Explore
  • Projects
  • Tools & AI
  • Network
  • Lab
  • Academy
FA/AR/EN
Sign inJoin
ISCDigital Humanities
Explore|Projects|Tools|Network|Lab|Academy|Home|About|Contact
FA/AR/EN

© 2026 ISC

Tools & AI

ISC Digital Humanities

Tools & AI

A research-tools directory for the ISC Digital Humanities platform — what's ready to use today, what's partially working, and what's planned.

12 capabilities ready to use now

Grouped by the kind of work each capability supports. Status reflects what is actually implemented and configured on this deployment, not a roadmap promise.

How a document moves through the platform

The tools below are stages of one workflow, not separate products. Each stage hands its output to the next.

  1. 01

    Upload

    Add a manuscript scan, PDF or image to a collection you can edit.

  2. 02

    Extract text

    Run OCR for printed sources or HTR for handwriting to turn the page image into text.

  3. 03

    Review & correct

    Check the extracted text against the page and fix it. Nothing is published until a human has read it.

  4. 04

    Translate & analyse

    Translate the corrected text, or run frequency and readability analysis over it.

Browse the collections

Extraction, translation and analysis run from the page of an individual document. Open a collection to reach one.

OCR — printed text recognition

Extract text from images using the selected free or paid mode. Preserve the original image and machine reading; review uncertain words and source-supported corrections.

ReadyExperimental on difficult pages
Setup details

Historical OCR benchmarks belong to the exact sample, model and version that produced them. They do not establish the accuracy of the currently selected API. Review this page against its image; uncertain readings remain drafts.

Get started

Document OCR (PDF & TIFF)

Extract text from PDF or TIFF pages. Availability, supported files and limits depend on the selected processing mode. Each page keeps its own result and status.

Ready
Get started

Handwritten Text Recognition

Handwritten and calligraphic text requires review against the image. No accuracy guarantee applies across scripts or historical hands; unsupported pages must be reported explicitly.

Disabled
Setup details

Handwriting recognition is switched off for this site.

Text Analysis & Mining

Real, local, deterministic computational analysis of transcribed texts — word frequency, n-grams, keyword-in-context concordance, and collocations, plus corpus-level statistics (token/type counts, lexical diversity). No AI provider or credential involved; open any item's Text Analysis tab to run it against its real transcription text.

Ready
Get started

Semantic Search

Vector/hybrid ranking layered on top of keyword search. One pgvector index serves both the AI Research Assistant's retrieval and Explore: once an item has been indexed — via the “Index for semantic search” action on the item itself — the first page of Explore's results is reordered by meaning rather than keyword score alone. Items that have not been indexed still surface normally through keyword search.

Disabled
Setup details

Semantic search is switched off for this site.

Machine Translation

On-demand translation of item transcriptions between English, Persian, and Arabic, through DeepL, a self-hosted LibreTranslate, or an AI-chat fallback. A translation is stored against the exact version of the original it was made from, so a later correction to the transcription marks the translation out of date instead of leaving a silent mismatch.

Ready
Get started

Natural Language Processing

Named-entity recognition, sentiment classification, and key-phrase extraction over item transcriptions in English, Persian, or Arabic. Analysis runs only on transcriptions a reviewer has verified, and where the configured analyser cannot handle the text's script it reports that it did not run rather than returning an empty or default result.

Ready
Get started

Knowledge Graph

A graph of relationships between researchers, institutions, projects, and labs. The interactive force-directed graph (pan, zoom, drag, hover, click-to-focus) is already built and renders live on the Network page from real database records — it just needs published ISC records with linked entities to show more than an empty graph.

Ready
Get started

GIS & Historical Mapping

Mapping historical places, routes, and spatial relationships. An interactive map (OpenStreetMap tiles, no API key required) is built and renders live on the Network page, plotting every published place with real coordinates from the real GeoJSON API — it just needs more published places with coordinates to show more than a handful of markers.

Ready
Get started

Data Visualization

Visual summaries of platform holdings. The Observatory page shows real, database-backed counts as both stat tiles and a bar chart — the same underlying figures, just two ways to read them.

Ready
Get started

AI Research Assistant

A conversational assistant that grounds its answers in retrieved ISC sources and says so when it has none. It needs the AI capability enabled and a provider configured. A deterministic offline demo mode is also available, always labelled as a demo and never presented as a live provider response.

Ready
Get started

Citation export

Plain-text, BibTeX, and CSL-JSON citations, generated from real record metadata — never fabricated authors or dates. Available on item, project, and dataset pages.

Ready
Get started

Annotations

Published Web Annotations attached to cultural items, viewable on the item detail page.

Ready
Get started

Transcriptions

Published transcriptions — human or machine-sourced — viewable on the item detail page, with source and confidence noted.

Ready
Get started

IIIF Manuscript Viewer

Page-by-page IIIF image viewing for digitized manuscripts with recorded digital representations.

Disabled
Setup details

The IIIF image service is switched off for this site.

NAN System (Needs and Answers Network)

An existing ISC digital platform being considered for regional deployment, per the ISC–ECOSF Draft Action Plan.

Ready
Get started

OCR tool (Noor Computer Research Center)

Per the DH Lab Framework pilot project list: an OCR tool built by Noor Computer Research Center, which the DH Lab plans to evaluate as part of OCR development for Miras-e Maktoob manuscripts.

Ready
Get started

These tools are the stages of one research workflow — from a page image to readable text, then translation and analysis. Every status above is read from this deployment's real configuration, the same capability matrix /api/health reports.