search-engine
Two-tier hybrid local search for Rust and the fsfs standalone CLI: fast first-pass results, then quality refinement.
Two-tier hybrid search for Rust: sub-millisecond initial results via potion-128M, quality-refined rankings in 150ms via MiniLM-L6-v2. Combines lexical (Tantivy BM25) and semantic (vector cosine) search with Reciprocal Rank Fusion. Progressive iterator API, f16 SIMD vector index, feature-gated compilation.
A high-performance document search engine built in Rust with WebAssembly support.
A high-performance document search engine built in Rust with WebAssembly support. Combines full-text search using FST (Finite State Transducers) with FSST compression for efficient storage and fast fuzzy matching capabilities.
Your files, found in seconds. Privately.
bHive reads what's inside your documents, photos and recordings — so you can find anything by describing it, or simply ask a question and get an answer drawn from your own files. All on your Mac. Nothing is ever uploaded.
Related contents:
Search engine aggregator with a comprehensive plugin/extension system.
Search aggregator that queries multiple engines and shows results in one place. You can add custom search engines, bang-command plugins, slot plugins (query-triggered panels above/below results or in the sidebar), and transports (custom HTTP fetch strategies like curl, FlareSolverr, or your own). The dream would be to eventually have a user made marketplace for plugins/engines.
camply, the campsite finder ⛺️, is a tool to help you book a campsite online. Finding reservations at sold out campgrounds can be tough. That's where camply comes in. It searches thousands of campgrounds across the USA world via the APIs of booking services like recreation.gov. It continuously checks for cancellations and availabilities to pop up - once a campsite becomes available, camply sends you a notification to book your spot!
Related contents:
This does for Documents what repo-browser does for repos. A local AI powered Document search engine.
DocuBrowse turns a messy pile of documents into something you can actually search. Point it at your files — PDFs, ebooks, Word docs, notes, whatever — and it builds a smart index that understands not just keywords, but meaning. Ask for "that contract about the lease renewal" and find it even if those exact words never appear. Click any result for an instant AI summary before you even open the file. PII aware and works with multiple document directories.
Research anything. Do anything.
Research at the speed of thought. The agentic research platform that plans, retrieves, and cites — so you can think faster.
Scira (Formerly MiniPerplx) is a minimalistic AI-powered search engine that helps you find information on the internet and cites it too. Powered by Vercel AI SDK!
The AI-native database. Built for the data your other databases can't touch.
Hybrid search. Local ML inference. Multimodal documents. One binary, zero glue code. Free to run in swarm mode, ready to scale with Antfly Cloud.
Related contents:
Le moteur de recherche qui respecte le web.
IBOU est un moteur de recherche conversationnel français. Conçu pour servir ceux qui cherchent l'information et ceux qui la créent.
Related contents:
A self-hosted, ad-free, privacy-respecting metasearch engine.
Get Google search results, but without any ads, JavaScript, AMP links, cookies, or IP address tracking. Easily deployable in one click as a Docker app, and customizable with a single config file. Quick and simple to implement as a primary search engine replacement on both desktop and mobile.
mini cli search engine for your docs, knowledge bases, meeting notes, whatever. Tracking current sota approaches while being all local.
An on-device search engine for everything you need to remember. Index your markdown notes, meeting transcripts, documentation, and knowledge bases. Search with keywords or natural language. Ideal for your agentic flows.
QMD combines BM25 full-text search, vector semantic search, and LLM re-ranking—all running locally via node-llama-cpp with GGUF models.
Related contents:
Codemogger is a code indexing library and MCP server for AI coding agents.
Code indexing library for AI coding agents. Parses source code with tree-sitter, chunks it into semantic units (functions, structs, classes, impl blocks), embeds them locally, and stores everything in a single SQLite file with vector + full-text search.
Related contents:
Web History on Steroids.
Blazing fast, content-based search for visited websites. Index the full content of web pages you visit, enabling meaningful search across your browsing hist.
Related contents:
A structural code search engine for AI agents. Provides sub-500ms file ranking across massive codebases without embeddings, vector databases, or external dependencies.
Mantic is an infrastructure layer designed to remove unnecessary context retrieval overhead for AI agents. It infers intent from file structure and metadata rather than brute-force reading content, enabling retrieval speeds faster than human reaction time.
WinFindr is an epic Windows file search tool.
WinFindr is a lightweight and easy-to-use Windows file and registry search tool, that can also search inside PDF files.
Related contents:
🔍 Tiny, full-text search engine for static websites built with Rust and Wasm.
Related contents:
YaCy is free software for your own search engine. Distributed Peer-to-Peer Web Search Engine and Intranet Search Appliance.
Related contents:
uBlacklist is a browser extension that filters Google Search results, available for Chrome, Firefox, and Safari.
Related contents:
an open source geocoder for openstreetmap data.
photon is an open source geocoder built for OpenStreetMap data. It is based on elasticsearch/OpenSearch - an efficient, powerful and highly scalable search platform.
Related contents:
Search engine for address. Only address.
Addok will index your address data and provide an HTTP API for full text search.
It is extensible with plugins, for example for geocoding CSV files.
Used in production by France administration, with around 26 millions addresses. In those servers, full France data is imported in about 15 min and it scales to around 2000 searches per second.
- Addok @ GitHub.
- Conteneurs Addok pour Docker avec les données de références diffusées par la Base Adresse Nationale :fr: @ GitHub.
Related contents:
Internet search engine for text-oriented websites. Indexing the small, old and weird web.
The aim of the project is to develop new and alternative discovery methods for the Internet. It's an experimental workshop as much as it is a public service, the overarching goal is to elevate the more human, non-commercial sides of the Internet.
Related contents:
Searloc is a fast, local randomizer for searXNG instances and for !!external bang users.
Related contents:
Static site search engine.
StaticSearch is a simple search engine you can add to any static website. It uses client-side JavaScript and JSON data files so there's no need for back-end server technologies or databases.
StaticSearch works with Publican but can be used on any static site built by any generator. It currently works best on English language sites, but most Western languages can be used.
AI-Powered Web Scraping & Data Enrichment. AI-powered web search with instant results and follow-up questions.
🔥 Blazing-fast AI search engine with real-time citations, streaming responses, and live data powered by Firecrawl
Related contents:
Decentralized search engine & automatized press reviews
- Explore the press with no middlemen between the newspapers and your web browser.
- Discover millions of results within seconds and explore the last ones in Firefox via this addon.
- Schedule searches, select your press review and export it in a few clicks.
Related contents:
Open Source, Distributed, Big Data Enterprise Search Engine.
Datafari is an open source enterprise search solution enriched with AI. It is the perfect product for anyone who needs to search and analyze its corporate data and documents, both within the content and the metadata. Plus, with its genAI modules, it allows to easily leverage mistral, openai, or local LLMs for your company data.
Filephish is a lightweight, user-friendly tool designed to help you quickly discover documents across the web by combining keywords, site-specific searches, and file type filters
A self-hosted BitTorrent indexer, DHT crawler, content classifier and torrent search engine with web UI, GraphQL API and Servarr stack integration.
the open source Meme Search Engine that's free and built to self-host locally with Python, Ruby, and Docker.
Scira (Formerly MiniPerplx) is a minimalistic AI-powered search engine that helps you find information on the internet. Powered by Vercel AI SDK! Search with models like Grok 2.0.
Stract is an open source web search engine hosted at stract.com targeted towards tinkerers and developers.
Unstructured Data Management Platform.
Open source file indexer, file search engine and data management and analytics powered by Elasticsearch
SEAL stands for: S earch E ngine A bstraction L ayer
The SEAL project is a PHP library designed to simplify the process of interacting with different search engines. It provides a straightforward interface that enables users to communicate with various search engines.
Full-text search for your desktop.
Recoll finds documents based on their contents as well as their file names. Recoll is based on the very capable Xapian search engine library, for which it provides a powerful text extraction layer and a complete, yet easy to use, Qt graphical interface.
thread-based email index, search and tagging.
Notmuch is a system for indexing, searching, reading, and tagging large collections of email messages in maildir or mh format. It uses the Xapian library to provide fast, full-text search with a convenient search syntax.
Search API, Search as a Service.
SeekStorm is an open-source, sub-millisecond full-text search library & multi-tenancy server implemented in Rust.
search config information for linux kernel modules.
Search for a package to see its download stats over time.
Visualize npm downloads in a beautiful chart, ready to be shared with your community.
🐤 Canary is modern Algolia DocSearch replacement.
Search & Ask AI across your docs(webpage), GitHub issues, and discussions.
batteries included search engine. Nixiesearch is a hybrid search engine that fine-tunes to your data.
Blazingly fast code search 🏎️ Deployed as a single Docker image 📦 Search million+ lines of code in your GitHub and GitLab repositories 🪄 MIT licensed ✅
Sourcebot is a fast code indexing and search tool for your codebases. It is built ontop of the zoekt indexer, originally authored by Han-Wen Nienhuys and now maintained by Sourcegraph.
An SQLite based, PHP-only fulltext search engine.
A full text search engine with tokenization, stemming, typo tolerance, filters and geo support based on only PHP and SQLite.
LLocalSearch is a completely locally running search aggregator using LLM Agents. The user can ask a question and the system will use a chain of LLMs to find the answer. The user can see the progress of the agents and the final answer. No OpenAI or Google API keys are needed.
Perplexica is an AI-powered search engine. It is an Open source alternative to Perplexity AI.
Perplexica is an open-source AI-powered searching tool or an AI-powered search engine that goes deep into the internet to find answers. Inspired by Perplexity AI, it's an open-source option that not just searches the web but understands your questions. It uses advanced machine learning algorithms like similarity searching and embeddings to refine results and provides clear answers with sources cited.
Open-source geocoding with OpenStreetMap data. Open Source search based on OpenStreetMap data.
Nominatim (from the Latin, 'by name') is a tool to search OpenStreetMap data by name and address (geocoding) and to generate synthetic addresses of OSM points (reverse geocoding). An instance with up-to-date data can be found at https://nominatim.openstreetmap.org. Nominatim is also used as one of the sources for the Search box on the OpenStreetMap home page.
Nominatim uses OpenStreetMap data to find locations on Earth by name and address (geocoding). It can also do the reverse, find an address for any location on the planet.
The Apache Tika™ toolkit detects and extracts metadata and text from over a thousand different file types (such as PPT, XLS, and PDF).
All of these file types can be parsed through a single interface, making Tika useful for search engine indexing, content analysis, translation, and much more.
Fess is very powerful and easily deployable Enterprise Search Server.
Fess is a very powerful and easily deployable Enterprise Search Server. You can quickly install and run Fess on any platform where you can run the Java Runtime Environment. Fess is provided under the Apache License 2.0.
Fess is based on OpenSearch/Elasticsearch, but knowledge/experience about OpenSearch/Elasticsearch is not required. Fess provides an easy to use Administration GUI to configure the system via your browser. Fess also contains a Crawler, which can crawl documents on a web server, file system, or Data Store (such as a CSV or database). Many file formats are supported including (but not limited to): Microsoft Office, PDF, and zip.
This website is the world most complete annotated UUID database, which happens to be an index of all the numbers between 0 and 2^128. Not all the numbers between 0 and 2^128 are RFC4122-compliant, but all are welcome in my database.
Meiliweb is a web-based administration panel that helps you store, organize and visualize data in your Meilisearch instances.
Search more with less.
Cloud-native search engine for observability. An open-source alternative to Datadog, Elasticsearch, Loki, and Tempo. Quickwit is the fastest search engine on cloud storage. It's the perfect fit for observability use cases.
Tantivy is a full-text search engine library inspired by Apache Lucene and written in Rust
Use the filters below to find projects by tag or counterpart. Information and frequently asked questions about this list can be found here.
Handy online tools for developers. Collection of handy online tools for developers, with great UX.
- Original: IT Tools @ GitHub.
- Maintained fork: IT Tools @ GitHub.
Related contents:
SearXNG is a free internet metasearch engine which aggregates results from more than 70 search services. Users are neither tracked nor profiled. Additionally, SearXNG can be used over Tor for online anonymity.
Related contents:
- SearX Search Provider @ Chrome Web Store.
- xSearch for Safari @ Apple Store.
- SearXNG : un meta moteur de recherche @ Linux things :penguin: :fr:.
- Episode 133 - No Google October @ Self Hosted.
- Private Internet Searches with SearXNG @ Jim's Garage.
- SearXNG — Privacy-focused metasearch engine for your homelab @ Akash Rajpurohit! 👋.
- Private Searching at Home: SearXNG Installation & Setup @ ServersatHome's YouTube.
Check Shortcuts of 101 Apps.
Designing keyboard shortcuts for your app can be a daunting task. High consistency with other tools is key to ensuring a minimal learning curve for your users. Shortcuts should also be conflict free with the system shortcuts to spare your users of rage when they accidentally try to print your app. Keycheck enables you to find the right shortcuts for your app in seconds.