@cyanheads/wikipedia-mcp-server
MCP serverSearchSearch Wikipedia, read summaries and full text, target sections, find nearby pages, list languages.
Available today. Use it from your connected AI after setup.
Needs your own MCP Auth Mode account. Credentials stay encrypted.
Connect ahel once, and every AI you use reads what you have installed.
Then ask your AI: use @cyanheads/wikipedia-mcp-server
From the project's README
As published by cyanheads/wikipedia-mcp-server in README.md.
Public Hosted Server: https://wikipedia.caseyjhand.com/mcp
Overview
Wikipedia content via the MediaWiki REST API and Action API. Search articles, read summaries or targeted sections, find geotagged pages near a coordinate, and list language editions from any MCP client. Runs as a stdio process, a local Streamable HTTP server, or the public hosted endpoint above.
Tools
| Tool | Description |
|---|---|
wikipedia_search_articles | Full-text search across Wikipedia, returning ranked results with plain-text snippets and page IDs. |
wikipedia_get_summary | Lead-section summary for any article — plain text, Wikidata QID, description, thumbnail URL, and page type. |
wikipedia_get_article | Full article or a targeted section as clean plain text, with section markers preserved. |
wikipedia_get_sections | Table of contents with section_index values for targeted section reads. |
wikipedia_search_nearby | Geotagged Wikipedia articles within a radius of a WGS 84 coordinate, sorted by distance. |
wikipedia_get_languages | All language editions available for an article, with titles and URLs. |
Capability reference
wikipedia_search_articles tool
- Free-text query, ranked by relevance; returns plain-text snippets (HTML stripped), page IDs, and word counts
limitcapped at 50;offsetpages further results — enrichmentnextOffsetsignals more remain and is passed back asoffsetlanguageselects any Wikipedia edition (defaulten)- Best when the exact article title is unknown, or to discover multiple articles on a topic
wikipedia_get_summary tool
- Returns the 2–4 paragraph lead extract, Wikidata QID (
wikibase_item), short description, and thumbnail URL page_typediscriminatesstandard/disambiguation/no-extract— ondisambiguation, re-query withwikipedia_search_articlesfor a more specific title- Redirect pages are followed automatically
- Right tool for most encyclopedic "what is X?" lookups; use
wikipedia_get_articlefor full depth
wikipedia_get_article tool
- Without
section_index: full article with== Section ==markers, unless it exceedsWIKIPEDIA_ARTICLE_OVERFLOW_BYTES(default 80,000 bytes) — then returns a section outline (truncated: true) pointing towikipedia_get_sectionsplus a targetedsection_indexread - With
section_index(fromwikipedia_get_sections): returns that section plus every nested subsection, each heading above its own body - Data tables are omitted from both paths — a section whose body is entirely a data table returns little beyond its heading; layout-only tables (multi-column lists, succession boxes) keep their content
- Page furniture — maintenance banners, sister-project and library-resource boxes, portal bars, spoken-article notices — is stripped; hatnotes are kept
- Redirect pages are followed automatically
wikipedia_get_sections tool
- Returns section titles, heading levels, hierarchical numbering (e.g.
"2.1"), andsection_indexvalues section_indexis the integer to pass towikipedia_get_articlefor a targeted read- Fails with
no_sectionson a stub or very short article — read it withwikipedia_get_articleinstead - Redirect pages are followed automatically
wikipedia_search_nearby tool
- Returns geotagged articles sorted ascending by distance, with coordinates and
distance_meters radius_meters: 10–10,000 (default 1000);limit: 1–500 (default 10) — no pagination pastlimit, so raise it or sweep narrower radii for full coverage- Only articles with a geographic coordinate in their Wikidata record are returned
- Enrichment
truncatedflags when more articles matched thanlimitallowed
wikipedia_get_languages tool
- Returns each edition's
language_code, tool-usableedition_code(can differ, e.g.gswvsals), article title, and URL - Pass
edition_code— notlanguage_code— as thelanguageparameter on other tools - Fails with
no_other_languageswhen the article has no translations - Redirect pages are followed automatically;
source_titlereports the resolved title
Features
Built on @cyanheads/mcp-ts-core: stdio and Streamable HTTP transports, pluggable auth (none / jwt / oauth), swappable storage (in-memory, filesystem, Supabase, Cloudflare KV/R2/D1), structured logging with optional OpenTelemetry tracing.
Wikipedia-specific:
- Dual API integration — MediaWiki REST API (
/api/rest_v1/) for summaries, Action API (/w/api.php) for search, full text, sections, geo search, and language links - Retry and backoff on all requests;
User-Agentheader per Wikimedia API policy - Both read paths render to the same plain-text shape —
== Heading ==markers, one list item per line — the full article from Action API extracts, a section from the parser's own HTML for that section. A section read additionally keeps code-sample indentation and the lists inside layout tables, neither of which the extract carries - Per-call
languageparameter on every tool — all Wikipedia language editions accessible in a single session - Language validation against a live edition registry built from the MediaWiki
action=sitematrixendpoint (cached 24h) — catches structurally valid but nonexistent editions before they cause timeouts
Agent-friendly output:
page_typeon summaries discriminatesstandard/disambiguation/no-extract— no string parsing neededwikibase_item(Wikidata QID) on summaries enables direct cross-referencing with wikidata-mcp-serversection_indexon table-of-contents entries links directly to the targeted-read parameter onwikipedia_get_article- Recovery hints on every error type — callers get actionable next steps (e.g., "use
wikipedia_search_articlesto find the correct title")
Getting started
Public Hosted Instance
A public instance is available at https://wikipedia.caseyjhand.com/mcp — no installation required. Point any MCP client at it via Streamable HTTP:
{
"mcpServers": {
"wikipedia-mcp-server": {
"type": "streamable-http",
"url": "https://wikipedia.caseyjhand.com/mcp"
}
}
}
Self-Hosted / Local
Add the following to your MCP client configuration file.
{
"mcpServers": {
"wikipedia-mcp-server": {
"type": "stdio",
"command": "bunx",
"args": ["@cyanheads/wikipedia-mcp-server@latest"],
"env": {
"MCP_TRANSPORT_TYPE": "stdio",
"MCP_LOG_LEVEL": "info"
}
}
}
}
Or with npx (no Bun required):
{
"mcpServers": {
"wikipedia-mcp-server": {
"type": "stdio",
"command": "npx",
"args": ["-y", "@cyanheads/wikipedia-mcp-server@latest"],
"env": {
"MCP_TRANSPORT_TYPE": "stdio",
"MCP_LOG_LEVEL": "info"
}
}
}
}
Or with Docker:
{
"mcpServers": {
"wikipedia-mcp-server": {
"type": "stdio",
"command": "docker",
"args": [
"run", "-i", "--rm",
"-e", "MCP_TRANSPORT_TYPE=stdio",
"ghcr.io/cyanheads/wikipedia-mcp-server:latest"
]
}
}
}
For Streamable HTTP, set the transport and start the server:
MCP_TRANSPORT_TYPE=http MCP_HTTP_PORT=3010 bun run start:http
# Server listens at http://localhost:3010/mcp
Prerequisites
- Bun v1.3.0 or higher (or Node.js v24+).
- No API keys required — Wikipedia's API is public.
Installation
- Clone the repository:
git clone https://github.com/cyanheads/wikipedia-mcp-server.git
- Navigate into the directory:
cd wikipedia-mcp-server
- Install dependencies:
bun install
- Configure environment (optional):
cp .env.example .env
# edit .env if you want to customize WIKIPEDIA_USER_AGENT or logging
Configuration
| Variable | Description | Default |
|---|---|---|
WIKIPEDIA_USER_AGENT | User-Agent header sent with every Wikimedia API request. Customize for your deployment. | wikipedia-mcp-server/0.2.0 (https://github.com/cyanheads/wikipedia-mcp-server) |
WIKIPEDIA_BASE_URL | Optional single-instance override. Unset (default): compose per-language hosts, language selects the edition per call. Set to a full base URL (e.g. a private MediaWiki mirror): route every call at that one fixed host — language no longer varies it. | (unset) |
WIKIPEDIA_ARTICLE_OVERFLOW_BYTES | Byte budget above which a full-article read (wikipedia_get_article without section_index) returns a section outline instead of the full text. Tuned for this domain — ordinary articles stay whole; only genuine mega-articles (World War II ~86 KB, United States ~94 KB) outline. Section-targeted reads are never affected. | 80000 |
MCP_TRANSPORT_TYPE | Transport: stdio or http. | stdio |
MCP_HTTP_PORT | Port for HTTP server. | 3010 |
MCP_SESSION_MODE | HTTP session mode: stateless, stateful, or auto (which resolves to stateful). The Docker image ships stateless. | auto |
MCP_AUTH_MODE | Auth mode: none, jwt, or oauth. | none |
MCP_LOG_LEVEL | Log level (RFC 5424). | info |
LOGS_DIR | Directory for log files (Node.js only). | <project-root>/logs |
OTEL_ENABLED | Enable OpenTelemetry instrumentation (spans, metrics, completion logs). | false |
See .env.example for the full list of optional overrides.
Running the server
Local development
-
Build and run:
# One-time build bun run rebuild # Run the built server bun run start:stdio # or bun run start:http -
Run checks and tests:
bun run devcheck # Lint, format, typecheck, security bun run test # Vitest test suite bun run lint:mcp # Validate MCP definitions against spec
Docker
docker build -t wikipedia-mcp-server .
docker run --rm -p 3010:3010 wikipedia-mcp-server
The Dockerfile defaults to HTTP transport, stateless session mode, and logs to /var/log/wikipedia-mcp-server. OpenTelemetry peer dependencies are installed by default — build with --build-arg OTEL_ENABLED=false to omit them.
Project structure
| Directory | Purpose |
|---|---|
src/index.ts | createApp() entry point — registers tools and inits the Wikipedia service. |
src/config | Server-specific environment variable parsing and validation with Zod. |
src/mcp-server/tools | Tool definitions (*.tool.ts) — one file per tool. |
src/services/wikipedia | WikipediaService — REST API + Action API client with retry/backoff and language validation. |
tests/ | Unit and integration tests mirroring src/. |
Development guide
See CLAUDE.md for development guidelines and architectural rules. The short version:
- Handlers throw, framework catches — no
try/catchin tool logic - Use
ctx.logfor request-scoped logging,ctx.statefor tenant-scoped storage - Register new tools in
src/mcp-server/tools/definitions/index.ts - Wrap external API calls: validate raw → normalize to domain type → return output schema; never fabricate missing fields
Contributing
Issues are welcome. Run checks and tests before submitting:
bun run devcheck
bun run test
License
Apache-2.0 — see LICENSE for details.
Advanced
- Delivery
- wikipedia-mcp-server MCP server → your ahel gateway (mcp.ahel.ai) → every connected AI client.
- Catalog kind
- mcp-server
- Gateway key
io-github-cyanheads-wikipedia-mcp-server- Source
- github.com/cyanheads/wikipedia-mcp-server
- Hosted endpoint
https://wikipedia.caseyjhand.com/mcp