
Official MCP RegistryListed
Islam West Africa Collection (IWAC)
Read-only access to the Islam West Africa Collection via Hugging Face datasets.
First seen 2 Oct 2026. Evidence as of 7 Oct 2026.
35
Tools
From an anonymous probe
1
Source listings
Each with its own history
45
Recorded changes
Since first seen
Tools
| Tool | Description | Behaviour |
|---|---|---|
| explore_corpus | Selection workbench: items, concordance, coverage, compare (requires comparison), attention, aliases, or paginated manifest/CSL-JSON/BibTeX exports. Exact filters intersect. Coverage measures archive availability. Exports identify this page and dataset snapshot; compare snapshot IDs across pages. Schema discovery: iwac://datasets/{subset}. | Read-only |
| fetch | Retrieve the full text and metadata of one IWAC item by an id returned from `search` (format '<category>:<number>', e.g. 'articles:28576'). Returns {id, title, text, url, metadata}: `text` is the item's OCR / abstract / transcription / description, `url` is the canonical islam.zmo.de link to cite, and `metadata` holds the remaining fields (author, date, country, newspaper, AI sentiment, …). Categories: articles, publications, references, documents, index, audiovisual, images. | Read-only |
| get_article | Get one article (by id): full metadata, the AI abstract (description_ai), AI sentiment, and OCR text, in 25k-char parts (follow `next_offset`). Pass a `keyword` to get ~2000-char excerpts around each match instead. | Read-only |
| get_audiovisual | Get one audiovisual record by id: full description and transcription (where one exists), creator/publishing channel, duration, medium, subjects, places, language, rights, source, and three distinct links: `url` (the IWAC page, the one to cite), `external_url` (where a harvested video plays) and `media_url` (a deposited file). `source_type` says which to expect. A long transcription comes in 25k-char parts (follow `next_offset`); `keyword` returns excerpts instead. | Read-only |
| get_collection_stats | Overall statistics for every IWAC subset, including `fulltext_coverage` — how many items in each subset actually carry searchable full text in this dataset. Read that before treating any keyword count as a full-text census. | Read-only |
| get_cooccurrence | Count pairs of metadata values co-tagged on the same items. Matrix/network views show descriptive associations, not causal links. Filters apply before pair counting; top_n caps the vocabulary. | Read-only |
| get_country_comparison | Compare article counts, newspaper counts, date ranges, and gpt-5-6-luna polarity across countries. | Read-only |
| get_document | Get one archival document (by id): full metadata, AI description, and OCR text, in 25k-char parts (follow `next_offset`). Pass a `keyword` to get ~2000-char excerpts around each match instead, useful for long documents. | Read-only |
| get_field_distribution | Rank exact metadata values, splitting pipe-separated tags. Reports missing metadata and optional coverage over time. Choose spatial for mentioned places, country for stored country tags, or subject/author/language. | Read-only |
| get_image | Get one photograph by id: title, photographer, capture date, place and coordinates, subjects, rights, the IIIF manifest, and the full-resolution `image_url`. The server returns URLs, not image bytes. | Read-only |
| get_index_entry | Get full details of an index entry by id (raw dataset columns, French names — Titre, Prénom, Coordonnées…). | Read-only |
| get_lexical_metrics | Text length, lexical richness and readability by country/newspaper/year. Metrics are precomputed; readability is French-specific and excludes other languages. Missing values are disclosed. Compare distributions as descriptive evidence. | Read-only |
| get_newspaper_stats | Per-newspaper article counts and date ranges. | Read-only |
| get_place_distribution | Place tags joined to authority coordinates. Separates publication countries from mentioned places; missing geocodes are counted explicitly. | Read-only |
| get_publication_fulltext | Full OCR text of a publication in 25k-char parts (follow `next_offset`), or ~2000-char excerpts around keyword matches (accent-insensitive; capped: see match_count vs excerpts_returned). | Read-only |
| get_reference | Full bibliographic record for one academic reference (by id), including the complete abstract (present for ~51% of references), subjects, DOI/URL, and host-work details (book, volume, issue, pages). | Read-only |
| get_semantic_map | Stable-hash sample of matching stored embeddings, projected with PCA. Reports eligible, invalid, missing and capped counts, seed and explained variance. Group counts describe the sample. Coordinates are chart-only; this is not the published UMAP landscape. Needs no API key. | Read-only |
| get_sentiment_distribution | Aggregate AI labels over a shared selection. model:"all" compares independent annotators and their coverage; model:"consensus" reads the stored panel majority. Choose compare_models and agreement_field for pairwise Cohen's kappa and quadratic weighted kappa, with explicit denominators. Labels are interpretations, not ground truth; subjectivity is weak evidence and its caveat must accompany findings. | Read-only |
| get_similar_items | Nearest stored embeddings by cosine similarity, without an API key. Scores rank reading candidates; no threshold proves a reprint. Returns source IDs and dates for comparison. Corpus-wide candidate generation belongs offline. | Read-only |
| get_temporal_distribution | Counts by Gregorian or Hijri year/month, or pooled lunar month. Group by country/newspaper. Partial dates and missing lunar dates are disclosed. normalize_by returns numerator counts and per-bucket denominators in the same country/outlet/date scope, removing thematic filters. Shares describe IWAC holdings. | Read-only |
| get_topic_distribution | Counts of offline LDA topic assignments, optionally over time. Topic labels are interpretive model outputs; exact topic_id filters select assignments rather than keyword matches. Reports unclassified coverage. | Read-only |
| list_audiovisual | List audiovisual materials, newest first (francophone web video from Burkina Faso, Togo and Benin; deposited Nigerian Hausa/Arabic recordings). Filter by country, publishing channel or `source_type`. | Read-only |
| list_locations | List lieux from the IWAC index, sorted by frequency (most-referenced first). The optional 'country' filter selects entries that APPEAR IN records from that country (mentioned-in, not located-in), ranked by collection-wide 'frequency' — so foreign and cross-border entries can appear. Nigeria returns none here (index frequency is computed from articles + publications + references, which have no Nigerian items — Nigeria is audiovisual only). | Read-only |
| list_periodicals | List the Islamic periodical/series titles in the publications subset, with issue counts and year ranges. Use the returned newspaper value as the `newspaper` filter on search_publications. | Read-only |
| list_persons | List personnes from the IWAC index, sorted by frequency (most-referenced first). The optional 'country' filter selects entries that APPEAR IN records from that country (mentioned-in, not located-in), ranked by collection-wide 'frequency' — so foreign and cross-border entries can appear. Nigeria returns none here (index frequency is computed from articles + publications + references, which have no Nigerian items — Nigeria is audiovisual only). | Read-only |
| list_subjects | List sujets from the IWAC index, sorted by frequency (most-referenced first). | Read-only |
| search | Search the Islam West Africa Collection across newspaper articles, Islamic publications, archival documents, academic references, audiovisual recordings, photographs, and the authority index (persons/places/organisations/events/subjects). Pass ONE concept or name — e.g. 'Tijaniyya', 'laïcité', 'Sheikh Gumi', 'pèlerinage'. Matching is accent- and case-insensitive; a multi-word query requires every word to appear somewhere in the item, so prefer a single concept per call. Write query strings and concept keywords in French for press/publication/document/index discovery even when the user's report language is not French. Academic references are multilingual, so try French and English title/abstract terms when relevant; metadata/filter labels remain French. Use the French transliteration of Islamic terms (Tabaski not 'Eid al-Adha', charia not 'sharia', Maouloud not 'Mawlid'). Returns {results:[{id,title,url,category}], ranking}; each result's `category` names its subset and the `ranking` field documents the ordering. Pass an id to `fetch` to read the full text. For filtered queries (by country, date, or newspaper) use the search_* tools instead. | Read-only |
| search_articles | Search IWAC newspaper articles by keyword (title + OCR + AI abstracts, French and English), country, newspaper, subject, and date range. Use French concept keywords regardless of the user's report language. Matching is accent- and case-insensitive. | Read-only |
| search_audiovisual | Search audiovisual materials by keyword and metadata: francophone web video from Burkina Faso, Togo and Benin (TV reports, association and campus recordings), plus deposited Nigerian Hausa/Arabic recordings. Keyword matches title, creator, publisher, subject, spatial, language, source, the item's own description (the richest text most of these items have) and its transcription where one exists. Each row says which population it is from (`source_type`) and carries either `external_url` (a video to watch) or `media_url` (a file), never both. | Read-only |
| search_by_sentiment | Filter articles by gpt-5-6-luna sentiment labels (accent/case-insensitive exact match). One model's reading, not a consensus — 4 other models scored the same articles and often disagree; get_sentiment_distribution with model:"all" shows by how much. `subjectivity` is much the weakest of the three scales, so treat a set selected on it as a lead to read rather than as a finding. | Read-only |
| search_documents | Search the small archival-documents subset (~26 items: Islamic association reports, flyers, project documents — mostly Burkina Faso). Use French concept keywords regardless of the user's report language. Most have OCR text and an AI description. Call with no arguments to list all. | Read-only |
| search_images | Search the IWAC photographs (30 items: mosques, radio stations, schools, signage and street scenes documented during fieldwork). Keyword matches title, creator, subject, place and the rare caption. Each result carries `image_url` (the full-resolution file), `coordinates` ('lat, lng' where known) and the canonical IWAC page. Call with no arguments to list all. Captions are almost never present, so prefer subject/place filters over keywords, or semantic_search_images when it is enabled. | Read-only |
| search_index | Search the IWAC authority index (persons, places, organisations, events, subjects) by name. Accent/case-insensitive. | Read-only |
| search_publications | Search Islamic publications (periodical issues, books). `keyword` matches title, subject, table of contents, and full OCR text (TOC hits come back as matching_toc_entries); use French concept keywords regardless of the user's report language. Filter by newspaper/series, subject, country and year. Use list_periodicals to discover series titles, and get_publication_fulltext for keyword excerpts from a single issue. | Read-only |
| search_references | Search academic references (journal articles, book chapters, theses, books, reports) by keyword and metadata. `keyword` is a single substring match over title + abstract, so search ONE term per call (combined terms like 'pèlerinage Mecque' miss results). References are multilingual: try French and English title/abstract keywords when relevant; metadata/filter values such as `reference_type` and `language` use French labels. Results include a short abstract snippet — use get_reference for the full abstract and bibliographic detail. | Read-only |
Change history
- search_references: input schema changed
- search_publications: input schema changed (+exact, +keyword_aliases, +keyword_mode)
- search_index: input schema changed
- search_images: input schema changed
- search_documents: input schema changed
- search_by_sentiment: input schema changed
- search_audiovisual: input schema changed
- search_articles: input schema changed (+exact, +keyword_aliases, +keyword_mode)
- search: input schema changed
- list_subjects: input schema changed
- list_persons: input schema changed
- list_locations: input schema changed
- list_audiovisual: input schema changed
- get_topic_distribution: input schema changed (+exact, +hijri_month, +hijri_year, +keyword_aliases, +keyword_mode)
- get_topic_distribution: description changed (+"Counts of offline" +"topic assignments, optionally over time. Topic labels" +"interpretive")
- get_temporal_distribution: input schema changed (+exact, +hijri_month, +hijri_year, +keyword_aliases, +keyword_mode, +normalize_by)
- get_temporal_distribution: description changed (+"or" +"year/month," +"pooled lunar month. Group by country/newspaper. Partial dates and missing lunar dates")
- get_similar_items: input schema changed
- get_similar_items: description changed (+"Nearest stored embeddings" +"similarity," +"an API key. Scores rank reading candidates;")
- get_sentiment_distribution: input schema changed (+agreement_field, +compare_models, +date_from, +date_to, +exact, +hijri_month, +hijri_year, +keyword, +keyword_aliases, +keyword_mode)
- get_sentiment_distribution: description changed (+"labels over" +"shared selection." +"compares independent annotators and their coverage; model:"consensus" reads")
- get_semantic_map: input schema changed (+exact, +hijri_month, +hijri_year, +keyword_aliases, +keyword_mode)
- get_semantic_map: description changed (+"Stable-hash sample" +"matching stored embeddings," +"with")
- get_reference: input schema changed
- get_publication_fulltext: input schema changed (+offset)
- get_publication_fulltext: description changed (+"publication in 25k-char parts (follow `next_offset`), or" +"capped:" -"publication, optionally returning")
- get_place_distribution: input schema changed (+exact, +hijri_month, +hijri_year, +keyword_aliases, +keyword_mode)
- get_place_distribution: description changed (+"Place tags" +"coordinates. Separates publication countries from mentioned places; missing geocodes" +"counted explicitly.")
- get_lexical_metrics: input schema changed (+exact, +hijri_month, +hijri_year, +keyword_aliases, +keyword_mode)
- get_lexical_metrics: description changed (+"Text length," +"readability" +"country/newspaper/year. Metrics are precomputed;")
- get_index_entry: input schema changed
- get_image: input schema changed
- get_field_distribution: input schema changed (+exact, +hijri_month, +hijri_year, +keyword_aliases, +keyword_mode)
- get_field_distribution: description changed (+"exact metadata values, splitting pipe-separated tags. Reports missing metadata and optional" +"over time. Choose spatial" +"mentioned places, country")
- get_document: input schema changed (+offset)
- get_document: description changed (+"text, in 25k-char parts (follow `next_offset`)." +"instead," -"text.")
- get_cooccurrence: input schema changed (+exact, +hijri_month, +hijri_year, +keyword_aliases, +keyword_mode)
- get_cooccurrence: description changed (+"Count pairs of metadata" +"co-tagged" +"same items. Matrix/network views show descriptive associations, not causal links. Filters apply before")
- get_collection_stats: description changed (-"public")
- get_audiovisual: input schema changed (+keyword, +offset)
- get_audiovisual: description changed (+"links:" +"A long transcription comes in 25k-char parts (follow `next_offset`); `keyword` returns excerpts instead." -"links —")
- get_article: input schema changed (+offset)
- get_article: description changed (+"text, in 25k-char parts (follow `next_offset`)." +"instead." -"text.")
- explore_corpus: tool added
- server instructions changed (+"IWAC" +"photographs" +"Nigeria")
| Source | Listing | First seen | Last seen | Versions |
|---|---|---|---|---|
| Official MCP Registry | io.github.fmadore/iwac-mcp-server | 2 Oct 2026 | 7 Oct 2026 | 1 |