Distributed P2P web search engine and intranet search appliance

4k stars 491 forks last commit first released

Actively maintained

Last commit 14 Jul 2026.

YaCy is a self-hosted search engine stack combining a web crawler, an index, and a web UI for searching and managing content. It can run as a standalone search portal, an intranet search appliance, or as part of a decentralized peer-to-peer network that exchanges index data for web search.

Key Features

  • Built-in web crawler with scheduling to keep indexes fresh
  • Search UI plus administration interface for configuring crawls, indexes, and peers
  • Peer-to-peer mode for sharing index data without relying on a central operator
  • Standalone mode for private, local-only search results from your own index
  • Intranet search use case with network scanning to discover HTTP, FTP, and SMB servers
  • HTTP-based interfaces with XML/JSON outputs for many pages and functions

Use Cases

  • Run a private search portal for a curated set of websites you crawl
  • Provide intranet search across internal web services and shared resources
  • Participate in a community-operated decentralized web search network

Limitations and Considerations

  • Precompiled packages may be less frequent; building from source is commonly recommended
  • Requires Java (11+) and can be resource-intensive depending on crawl and index size

YaCy is suited to organizations and individuals who want control over crawling and indexing, and who prefer privacy-aware search without dependence on a centralized search provider. Its flexible modes make it useful both for private indexing and for distributed web search participation.

Categories:

Tags:

Tech Stack:

Share:

Similar to YaCy

SearXNG logo

SearXNG

Privacy-focused metasearch engine for aggregating web results

34.5k
3.2k
Last commit

SearXNG is a privacy-respecting metasearch engine that aggregates results from many search services without tracking or profiling users.

AGPL-3.0Actively maintained
Alternative to:
Google Search logo
Google Search
+6
Manticore Search logo

Manticore Search

Fast open-source search database with SQL and JSON APIs

11.9k
634
Last commit

Manticore Search is a fast open-source search database for full-text, faceted, and vector search with SQL (MySQL protocol) and HTTP JSON APIs.

GPL-3.0Actively maintained
Alternative to:
Manticore Search logo
Manticore Search
+15
OpenSearch logo

OpenSearch

Distributed search and analytics engine with a RESTful API

13.4k
2.8k
Last commit

OpenSearch is an Apache 2.0 open source distributed search and analytics engine for indexing, querying, and analyzing large-scale data with REST APIs.

Apache-2.0Actively maintained
Alternative to:
Amazon OpenSearch Service logo
Amazon OpenSearch Service
+19
wanderer logo

wanderer

Self-hosted trail database for planning and searching GPS tracks

3.8k
181
Last commit

Self-hosted trail catalog to plan routes, upload GPX/GPS tracks, add metadata, and search, filter, and share trails with optional federation.

AGPL-3.0Actively maintained
Alternative to:
Ride with GPS logo
Ride with GPS
+2
Feedbin logo

Feedbin

Web-based RSS reader with search, full-text and automation

3.8k
289
Last commit

Feedbin is a web-based RSS reader for organizing and reading feeds with full-text extraction, powerful search, filtering actions, and a REST-like API for clients.

MITActively maintained
Alternative to:
Feedly logo
Feedly
+19
Fess logo

Fess

Enterprise full-text search server with built-in crawler and admin UI

1.1k
173
Last commit

Fess is an open-source enterprise search server with a built-in crawler, web-based administration, and OpenSearch/Elasticsearch-backed full-text search.

Apache-2.0Actively maintained
Alternative to:
Glean logo
Glean
+9