Showing 75 open source projects for "keyword"

View related business solutions
  • $300 Free Credits to Build on Google Cloud Icon
    $300 Free Credits to Build on Google Cloud

    New customers can spin up VMs, build with AI, and query data at no cost.

    Put your $300 in credit toward real workloads, then keep building with free monthly usage for 20+ products. No commitment and no charge until you upgrade.
    Start Free
  • Save Up to 91% on Cloud Compute With Spot VMs Icon
    Save Up to 91% on Cloud Compute With Spot VMs

    Automatic sustained-use discounts. One free VM per month. No negotiation needed.

    Run batch jobs at 60-91% off with Spot VMs. Long-running workloads get automatic discounts with sustained use.
    Try Free
  • 1
    MediaCrawler

    MediaCrawler

    Self-media platform crawler

    ...The project uses Playwright to automate browser login and preserve authenticated sessions. It obtains required request signatures from the active browser context, avoiding complex JavaScript reverse engineering. Users can search by keyword, crawl specific posts, collect nested comments, and retrieve creator profiles. It also supports cached login states, proxy pools, and comment word-cloud generation. The repository is intended for learning and research, with clear restrictions against illegal, large-scale, or commercial scraping.
    Downloads: 9 This Week
    Last Update:
    See Project
  • 2
    MoneyPrinterTurbo

    MoneyPrinterTurbo

    Generate short videos with one click using AI LLM

    MoneyPrinterTurbo is an AI-driven tool that enables users to generate high-definition short videos with minimal input. By providing a topic or keyword, the system automatically creates video scripts, sources relevant media assets, adds subtitles, and incorporates background music, resulting in a polished video ready for distribution.
    Downloads: 575 This Week
    Last Update:
    See Project
  • 3
    CineCLI

    CineCLI

    CineCLI is a cross-platform command-line movie browser

    ...It connects to popular online movie databases to fetch metadata such as titles, release dates, ratings, genres, casts, posters, and plot summaries, presenting all of that in a concise, text-friendly format suitable for terminals or scripts. Users can search by keyword, year, or exact title and then drill into detailed views for individual films, making it useful for creating watchlists, learning more about films before watching, or integrating movie lookup into shell workflows. CineCLI also supports paginated results and filters so users can navigate large search outputs without overwhelming their screens. ...
    Downloads: 7 This Week
    Last Update:
    See Project
  • 4
    Semantra

    Semantra

    Multi-tool for semantic search

    Semantra is an open-source semantic search tool designed to help users explore large collections of documents by meaning rather than simple keyword matching. The software analyzes text and PDF documents stored locally and creates embeddings that allow queries to retrieve results based on conceptual similarity. It is primarily intended for individuals who need to extract insights from large document collections, including researchers, journalists, students, and historians. The system runs from the command line and automatically launches a local web interface where users can perform interactive searches and examine document passages related to a query. ...
    Downloads: 5 This Week
    Last Update:
    See Project
  • Veeam Data Platform v13.1 - Get Your Free Trial Icon
    Veeam Data Platform v13.1 - Get Your Free Trial

    Secure by design, portable by default. Recover clean, fast, anywhere. Start a free trial.

    Try Veeam Data Platform today. Experience the unified platform that's secure by design, portable by default, and proven to recover clean, fast, and anywhere.
    Try it Free
  • 5
    TrendRadar

    TrendRadar

    AI-driven public opinion trend monitor with multi-platform aggregation

    TrendRadar is an AI-powered trend and hotspot tracking system that aggregates information from dozens of news, social, and content platforms to help users cut through information overload and focus on what matters. It automatically crawls and monitors trends across 30+ sources with smart filtering, keyword triggers, sentiment analysis, and natural language summarization to give actionable insights. The tool supports multiple alert modes—such as daily summaries, incremental change monitoring, and current rankings—and can push notifications through messaging platforms like Telegram, Slack, WeChat, DingTalk, and email. Users can deploy it quickly via Python and GitHub Actions, and it also supports RSS feeds and Docker deployment for flexible integration. ...
    Downloads: 2 This Week
    Last Update:
    See Project
  • 6
    LocalGPT

    LocalGPT

    Chat with your documents on your local device using GPT models

    LocalGPT is a private, on-premises document intelligence platform for questioning, summarizing, and analyzing files with locally hosted language models. Its data remains on the user’s machine, making it suitable for confidential or offline workflows. The retrieval system combines semantic similarity, keyword matching, late chunking, contextual enrichment, and sentence-level pruning. A smart router chooses between retrieval-augmented generation and direct model responses for each query. An independent verification pass is intended to improve answer reliability before results are returned. The platform works with Ollama models, offers a browser interface and API, and can run through local or Docker-based setups. ...
    Downloads: 0 This Week
    Last Update:
    See Project
  • 7
    DocuBrowse

    DocuBrowse

    This does for Documents what repo-browser does for repos

    DocuBrowse is a local AI-powered document search engine for personal or professional file collections. It indexes documents such as PDFs, Word files, spreadsheets, presentations, ebooks, HTML, text, and Markdown. The system combines SQLite FTS5 keyword search with local semantic embeddings for meaning-based retrieval. Users can click search results to generate an AI synopsis before opening the original file. It runs entirely on the user’s machine with Ollama models, so it does not require accounts, API keys, internet access, or per-query cloud costs. It also includes PII detection, multiple document directories, tag filtering, duplicate cleanup, and packaged installers for Linux, Windows, and macOS.
    Downloads: 0 This Week
    Last Update:
    See Project
  • 8
    SeaGOAT

    SeaGOAT

    local-first semantic code search engine

    ...By combining vector search with tools like ripgrep, SeaGOAT provides a hybrid approach that supports both natural language queries and precise keyword matching in source files. It is built primarily in Python and is intended to work on common operating systems such as Linux, macOS, and Windows.
    Downloads: 1 This Week
    Last Update:
    See Project
  • 9
    SEO Machine

    SEO Machine

    A specialized Claude Code workspace for creating long-form

    ...It integrates research, writing, analysis, and optimization into a single pipeline, allowing users to produce high-quality articles tailored to search engine performance. The system uses specialized commands and agents to perform tasks such as keyword research, competitor analysis, content drafting, and optimization. It incorporates real data sources like Google Analytics and Search Console to guide decision-making and improve content effectiveness. The architecture emphasizes context-awareness, using brand voice, style guides, and keyword strategies to maintain consistency across outputs. ...
    Downloads: 0 This Week
    Last Update:
    See Project
  • MongoDB Atlas runs apps anywhere Icon
    MongoDB Atlas runs apps anywhere

    Deploy in 115+ regions with the modern database for every enterprise.

    MongoDB Atlas gives you the freedom to build and run modern applications anywhere—across AWS, Azure, and Google Cloud. With global availability in over 115 regions, Atlas lets you deploy close to your users, meet compliance needs, and scale with confidence across any geography.
    Start Free
  • 10
    NLP

    NLP

    Open source NLP guide with models, methods, and real use cases

    ...Its covers core NLP concepts such as text representation, feature extraction, and model evaluation, alongside hands-on implementations using tools like Word2Vec, TF-IDF, and FastText. It also introduces topic modeling with LDA, keyword extraction techniques, and document similarity methods. NLP extends into real-world applications, including sentiment analysis and text classification, helping readers connect concepts to use cases. Designed for accessibility, the project evolves over time, allowing updates and improvements as NLP techniques advance. It reflects a practical approach to learning, where readers can explore code, experiment with models, and build foundational skills in machine learning-driven language processing.
    Downloads: 0 This Week
    Last Update:
    See Project
  • 11
    geo-seo-claude

    geo-seo-claude

    GEO-first SEO skill for Claude Code

    ...It leverages AI to generate location-specific content tailored to different regions, allowing users to scale SEO efforts across multiple cities or markets without manual content creation. The system focuses on producing structured and keyword-optimized pages that align with search engine ranking factors, including localized relevance and semantic context. It is particularly useful for agencies, marketers, and businesses that need to manage large volumes of localized landing pages efficiently. Geo SEO Claude can integrate with existing content pipelines, enabling automated generation and deployment of SEO assets. ...
    Downloads: 1 This Week
    Last Update:
    See Project
  • 12
    Scweet

    Scweet

    Scrape tweets, profiles, followers and following from Twitter/X

    ...Instead of depending on deprecated unauthenticated scraping methods, it works by using X’s web GraphQL API together with authenticated browser cookies, which gives it a more current and practical approach for data extraction. The project supports a broad set of collection patterns, including searches by keyword, hashtag, user, date range, engagement thresholds, language, and location, making it useful for research, monitoring, and data gathering workflows. It is built for both local use and higher-volume runs, with support for proxies, dedicated accounts, and multi-account cookie handling to improve reliability at scale. Scweet also includes asynchronous method variants, a command-line interface, automatic credential persistence in a local database, etc.
    Downloads: 0 This Week
    Last Update:
    See Project
  • 13
    cloud_enum

    cloud_enum

    Multi-cloud OSINT tool for discovering public cloud resources

    cloud_enum is an open source reconnaissance and OSINT tool designed to discover publicly accessible cloud resources across major cloud providers. It focuses on enumerating assets in Amazon Web Services (AWS), Microsoft Azure, and Google Cloud Platform using keyword-based discovery techniques. It works by taking user-provided keywords and generating variations through mutation wordlists, then testing these combinations against common cloud service naming patterns. cloud_enum performs both HTTP probing and DNS lookups to identify resources such as storage buckets, cloud applications, and databases that may be exposed or accessible. cloud_enum uses concurrent processing to speed up scanning, enabling efficient enumeration of large numbers of possible resource names. ...
    Downloads: 0 This Week
    Last Update:
    See Project
  • 14
    Wfuzz

    Wfuzz

    Web application fuzzer

    Wfuzz provides a framework to automate web applications security assessments and could help you to secure your web applications by finding and exploiting web application vulnerabilities. Wfuzz it is based on a simple concept: it replaces any reference to the FUZZ keyword by the value of a given payload. A payload in Wfuzz is a source of data. This simple concept allows any input to be injected in any field of an HTTP request, allowing to perform complex web security attacks in different web application components such as: parameters, authentication, forms, directories/files, headers, etc.
    Downloads: 16 This Week
    Last Update:
    See Project
  • 15
    PaperAI

    PaperAI

    Semantic search and workflows for medical/scientific papers

    PaperAI is an open-source framework for searching and analyzing scientific papers, particularly useful for researchers looking to extract insights from large-scale document collections.
    Downloads: 0 This Week
    Last Update:
    See Project
  • 16
    AIOHTTP

    AIOHTTP

    Asynchronous HTTP client/server framework for asyncio and Python

    ...If you were constructing the URL by hand, this data would be given as key/value pairs in the URL after a question mark, e.g. httpbin.org/get?key=val. Requests allows you to provide these arguments as a dict, using the params keyword argument. aiohttp internally performs URL canonicalization before sending request.
    Downloads: 13 This Week
    Last Update:
    See Project
  • 17
    dirsearch

    dirsearch

    Web path scanner

    An advanced command-line tool designed to brute force directories and files in webservers, AKA web path scanner. Wordlist is a text file, each line is a path. About extensions, unlike other tools, dirsearch only replaces the %EXT% keyword with extensions from -e flag. For wordlists without %EXT% (like SecLists), -f | --force-extensions switch is required to append extensions to every word in wordlist, as well as the /. To use multiple wordlists, you can separate your wordlists with commas. Example: wordlist1.txt,wordlist2.txt. Default values for dirsearch flags can be edited in the configuration file: default.conf. ...
    Downloads: 8 This Week
    Last Update:
    See Project
  • 18
    Kirara AI

    Kirara AI

    DIY multimodal AI chatbot

    Kirara AI is a chatbot framework designed to connect large language models with mainstream chat platforms. It supports multi-turn conversation, persona presets, keyword-triggered replies, administrator commands, and conditional automation. The project can work with model providers such as OpenAI, DeepSeek, Claude, Gemini, Qwen, Mistral, Kimi, Minimax, and others. It also supports image generation models, voice replies, cross-platform message sending, and custom workflows. A WebUI helps users manage models, workflows, plugins, and platform settings without editing every detail manually. ...
    Downloads: 1 This Week
    Last Update:
    See Project
  • 19
    WeKnora

    WeKnora

    LLM framework for document understanding and semantic retrieval

    WeKnora is an open source framework developed for deep document understanding and semantic information retrieval using large language models. It focuses on analyzing complex and heterogeneous documents by combining multiple processing stages such as multimodal document parsing, vector indexing, and intelligent retrieval. It follows the Retrieval-Augmented Generation (RAG) paradigm, where relevant document segments are retrieved and used by language models to generate accurate, context-aware...
    Downloads: 5 This Week
    Last Update:
    See Project
  • 20
    Hindsight

    Hindsight

    Hindsight: Agent Memory That Learns

    Hindsight is an advanced, open-source memory system for AI agents designed to enable long-term learning, reasoning, and consistency across interactions by treating memory as a first-class component of intelligence rather than a simple retrieval layer. It addresses one of the core limitations of modern AI agents, which is their inability to retain and meaningfully use past experiences over time, by introducing a structured, biomimetic memory architecture inspired by how human memory works....
    Downloads: 4 This Week
    Last Update:
    See Project
  • 21
    Claude Cognitive

    Claude Cognitive

    Persistent context and multi-instance coordination

    ...It introduces an attention-based context router that prioritizes files and content relevant to the current development discussion — tagging them as HOT, WARM, or COLD based on recency and keyword activation — so Claude Code doesn’t waste token budget rereading irrelevant code. This context routing dramatically reduces redundant token usage and accelerates large codebase interactions by focusing only on what truly matters to the current task. Additionally, Claude-Cognitive includes a pool coordinator to share state across multiple Claude Code instances, preserving what’s been learned or completed and preventing repetitive debugging or redundant exploration.
    Downloads: 1 This Week
    Last Update:
    See Project
  • 22
    Paperless-AI

    Paperless-AI

    AI-powered document analysis and tagging for Paperless-ngx

    ...A key capability is its use of retrieval-augmented generation, which enables semantic search and natural language interaction across an entire document archive. Users can ask contextual questions about their files and receive precise answers based on full document understanding rather than simple keyword matching. Paperless-AI also includes a web interface for manual review and tagging, allowing greater control when handling sensitive or complex documents.
    Downloads: 2 This Week
    Last Update:
    See Project
  • 23
    newspaper4k

    newspaper4k

    Python library for scraping and analyzing online news articles easily

    Newspaper4k is a Python library designed for extracting, processing, and analyzing news articles from websites. It is a continuation and active fork of the original newspaper3k library, which had stopped receiving updates, with the goal of keeping the ecosystem maintained while adding improvements and bug fixes. It provides developers with tools to automatically download web pages, extract the main article content, and collect associated metadata such as titles, authors, images, and...
    Downloads: 2 This Week
    Last Update:
    See Project
  • 24
    ZeusDB Vector Database

    ZeusDB Vector Database

    Blazing-fast vector DB with similarity search and metadata filtering

    ...Developers get multiple ingestion paths—batch, streaming, and upsert—making it easy to keep embeddings synchronized as content changes. Hybrid search is a core design goal, allowing you to mix vector, keyword, and filter queries in a single request for practical relevance. Observability and safety round out the system, with metrics, tracing, and guardrails to manage recalls, deletions, and privacy-sensitive data at scale.
    Downloads: 2 This Week
    Last Update:
    See Project
  • 25
    Mangum

    Mangum

    AWS Lambda support for ASGI applications

    ...The heart of Mangum is the adapter class. It is a configurable wrapper that allows any ASGI application (or framework) to run in an AWS Lambda deployment. The adapter accepts a number of keyword arguments to configure settings related to logging, HTTP responses, ASGI lifespan, and API Gateway configuration.
    Downloads: 2 This Week
    Last Update:
    See Project
  • Previous
  • You're on page 1
  • 2
  • 3
  • Next