Showing 282 open source projects for "index data"

View related business solutions
  • Ship Agents Faster Icon
    Ship Agents Faster

    Transform your applications and workflows into powerful agentic systems at global scale.

    Gemini Enterprise Agent Platform lets you rapidly build, scale, govern and optimize production-ready agents grounded in your organization's data. The platform enables developers to build custom or pre-built agents for virtually any use case. New customers get $300 in free credits.
    Start Free
  • Custom VMs From 1 to 96 vCPUs With 99.95% Uptime Icon
    Custom VMs From 1 to 96 vCPUs With 99.95% Uptime

    General-purpose, compute-optimized, or GPU/TPU-accelerated. Built to your exact specs.

    Live migration and automatic failover keep workloads online through maintenance. One free e2-micro VM every month.
    Start Free
  • 1
    Scout Elasticsearch Driver

    Scout Elasticsearch Driver

    Offers advanced functionality for searching data in Elasticsearch

    This package offers advanced functionality for searching and filtering data in Elasticsearch.
    Downloads: 0 This Week
    Last Update:
    See Project
  • 2
    DrQA

    DrQA

    Reading Wikipedia to Answer Open-Domain Questions

    ...The retriever relies on classic IR features (like TF-IDF and n-gram statistics) to remain lightweight and scalable to millions of documents. The reader is a neural model trained on supervised QA data to estimate start and end positions within a paragraph, and it can be adapted to new domains through fine-tuning or distant supervision. The repository includes scripts to build the Wikipedia index, train the reader, and evaluate end-to-end performance. DrQA popularized a practical recipe for combining IR and neural reading, and it remains a strong baseline for open-domain QA research and production prototypes.
    Downloads: 0 This Week
    Last Update:
    See Project
  • 3
    MachineLearningStocks

    MachineLearningStocks

    Using python and scikit-learn to make stock predictions

    MachineLearningStocks is a Python-based template project that demonstrates how machine learning can be applied to predicting stock market performance. The project provides a structured workflow that collects financial data, processes features, trains predictive models, and evaluates trading strategies. Using libraries such as pandas and scikit-learn, the repository shows how historical financial indicators can be transformed into machine learning features. The model attempts to predict whether specific stocks will outperform a benchmark index such as the S&P 500. ...
    Downloads: 0 This Week
    Last Update:
    See Project
  • 4

    DEPRECATED - KVFinder

    Cavity Detection PyMOL plugin

    ...Please read and cite the original paper ParKVFinder: A thread-level parallel approach in biomolecular cavity detection (10.1016/j.softx.2020.100606). [pyKVFinder] pyKVFinder is available in this Python Package Index (PyPI) repository, https://pypi.org/project/pyKVFinder and this GitHub repository, https://github.com/LBC-LNBio/pyKVFinder. Please read and cite the original paper pyKVFinder: an efficient and integrable Python package for biomolecular cavity detection and characterization in data science (10.1186/s12859-021-04519-4).
    Downloads: 0 This Week
    Last Update:
    See Project
  • Earn up to 16% annual interest with Nexo. Icon
    Earn up to 16% annual interest with Nexo.

    Access competitive interest rates on your digital assets.

    Generate interest, borrow against your crypto, and trade a range of cryptocurrencies — all in one platform. Geographic restrictions, eligibility, and terms apply.
    Get started with Nexo.
  • 5
    CyC2018.github.io

    CyC2018.github.io

    Personal knowledge site built with GitHub Pages

    This is a personal knowledge site built with GitHub Pages, organizing computer science notes and study materials in a browsable format. It aggregates content across algorithms, data structures, networking, operating systems, databases, and interview preparation into a coherent index. The presentation emphasizes succinct explanations paired with diagrams, tables, or code snippets to speed recall. Because the site is generated from a repository, it benefits from issue tracking, pull requests, and version history, making it easy to update and maintain. ...
    Downloads: 0 This Week
    Last Update:
    See Project
  • 6
    fastNLP

    fastNLP

    fastNLP: A Modularized and Extensible NLP Framework

    fastNLP is a lightweight framework for natural language processing (NLP), the goal is to quickly implement NLP tasks and build complex models. A unified Tabular data container simplifies the data preprocessing process. Built-in Loader and Pipe for multiple datasets, eliminating the need for preprocessing code. Various convenient NLP tools, such as Embedding loading (including ELMo and BERT), intermediate data cache, etc.. Provide a variety of neural network components and recurrence models...
    Downloads: 0 This Week
    Last Update:
    See Project
  • 7
    StructPie

    StructPie

    A set of C libraries to implement data structures and algorithms

    ...In the "hash_table" directory, the hash table implementation uses linked lists. While in "HashBSTree" directory, a hash table with a binary tree in each index is implemented for faster lookup in large data. The stack, the tree and the hash table accept int, float and char* data type To look at the python library : https://github.com/mnoorfawi/struct-pie
    Downloads: 0 This Week
    Last Update:
    See Project
  • 8

    DBRECOVER for MySQL

    repair corrupt MySQL database recover dropped table and database

    ...In general terms, the tool works by extracting index pages from the monolithic ibdata file(s) and/or from independent InnoDB tablespace files (.ibd files where innodb_file_per_table is in use). Where applicable, blob pages are extracted to a separate subdirectory. After MySQL drops a table the data is still on the media for a while. So you can fetch records and rebuild a table with DBRECOVER.
    Downloads: 1 This Week
    Last Update:
    See Project
  • 9
    JuliaDB.jl

    JuliaDB.jl

    Parallel analytical database in pure Julia

    JuliaDB is a package for working with large persistent data set. JuliaDB provides distributed table and array datastructures with convenient functions to load data from CSV. JuliaDB is Julia all the way down. This means queries can be composed with Julia code that may use a vast ecosystem of packages.
    Downloads: 0 This Week
    Last Update:
    See Project
  • Train ML Models With SQL You Already Know Icon
    Train ML Models With SQL You Already Know

    BigQuery automates data prep, analysis, and predictions with built-in AI assistance.

    Build and deploy ML models using familiar SQL. Automate data prep with built-in Gemini. Query 1 TB and store 10 GB free monthly.
    Start Free
  • 10
    hordes

    hordes

    WordPress Plug curates list of links with titles icons and categories.

    Creates a front side submitted form that allows logged in user to add links or bookmarks of their favorite websites. Just copy and paste your address into the link text box and add then a title. Hordes will do the rest. There are many options including the ability to show all the data about your link, or to only show the title and the link which is handy to show more links on the page when you know what they are by only looking at the title and the link. If you need to search for...
    Downloads: 0 This Week
    Last Update:
    See Project
  • 11
    nodebook

    nodebook

    Multi-Lang Web REPL

    Useful to practice algorithms and data structures for coding interviews. Nodebook is an in-browser REPL supporting many programming languages. Code's on the left, Console's on the right. Click "Run" or press Ctrl+Enter or Cmd+Enter to run your code. Code is automatically persisted on the file system. You can also use Nodebook directly on the command line, running your notebooks upon change.
    Downloads: 0 This Week
    Last Update:
    See Project
  • 12
    Happy Java Library

    Happy Java Library

    Multilock, Collections, Controllers, Delegates, Generators, Streams

    Helps to develop and test event-based multi-threaded Java application. Because of method called as API-Evolution the Happy Java Library is fully downward compatible. The library contains following functionality: MultiLock, Parallel loops, Collections, Controllers, Generators, Delegates, Streams.
    Downloads: 0 This Week
    Last Update:
    See Project
  • 13
    TP-COBOL-DEBUGGER

    TP-COBOL-DEBUGGER

    A COBOL debugger for GnuCOBOL/OpenCOBOL written in GnuCOBOL

    A COBOL debugger for GnuCOBOL written in GnuCOBOL. Works with both current GnuCOBOL and old GnuCOBOL/OpenCOBOL 1.1; could be used for other vendors with slightly modifications, too. Take a look at https://gnucobol.altervista.org/tp-cobol-debugger/
    Downloads: 2 This Week
    Last Update:
    See Project
  • 14
    Mahuta

    Mahuta

    IPFS Storage service with search capability

    Mahuta (formerly known as IPFS-Store) is a library to aggregate and consolidate files or documents stored by your application on the IPFS network. It provides a solution to collect, store, index, cache and search IPFS data handled by your system in a convenient way.
    Downloads: 0 This Week
    Last Update:
    See Project
  • 15

    OLiA

    OWL/DL ontologies for linguistic annotations

    MOVED TO https://github.com/acoli-repo/olia. The Ontologies of Linguistic Annotations (OLiA) provide an OWL/DL taxonomy of data categories as a reference for linguistic annotation (OLiA Reference Model), plus OWL/DL models for a large number of annotation schemes (OLiA Annotation Models) and their relationship to reference data categories (OLiA Linking Models). The OLiA Reference Model itself is linked to community-maintained repositories such as GOLD (http://linguistics-ontology.org/) and ISOcat (http://www.isocat.org) The OLiA ontologies were originally developed as part of an infrastructure for the sustainable maintenance of linguistic resources (http://www.sfb441.uni-tuebingen.de/c2/index-engl.html), their fields of application include the formalization of annotation schemes, concept-based querying over heterogeneously annotated corpora, and the development of interoperable NLP pipelines.
    Downloads: 1 This Week
    Last Update:
    See Project
  • 16
    Mirage

    Mirage

    GUI for simplifying Elasticsearch Query DSL

    ...It offers a blocks-based GUI for composing Elasticsearch queries and comes with an on-the-fly transformer to show the corresponding JSON query API of Elasticsearch. Queries can be saved for later reuse. They can also be captured and shared by copying the URL. Mirage works with any Elasticsearch 2. x, 5. x, 6. x, and 7.x index currently. Below is the roadmap for query support. Every app in appbase.io has a query explorer view, which uses mirage. Mirage can be used along with ⊞ DejaVu to browse data and perform CRUD operations inside an Elasticsearch index.
    Downloads: 0 This Week
    Last Update:
    See Project
  • 17
    Elasticquent

    Elasticquent

    Maps Laravel Eloquent models to Elasticsearch types

    ...Elasticquent makes working with Elasticsearch and Eloquent models easier by mapping them to Elasticsearch types. You can use the default settings or define how Elasticsearch should index and search your Eloquent models right in the model. Elasticquent uses the official Elasticsearch PHP API. To get started, you should have a basic knowledge of how Elasticsearch works (indexes, types, mappings, etc). When using a database, Eloquent models are populated from data read from a database table. With Elasticquent, models are populated by data indexed in Elasticsearch. ...
    Downloads: 0 This Week
    Last Update:
    See Project
  • 18

    owl-indexer

    Full-text index generator of OWL literals

    owl-indexer generates full text index of OWL literal via either Apache Lucene or Sphinx Search. It is based on OWL API (https://github.com/owlcs/owlapi).
    Downloads: 0 This Week
    Last Update:
    See Project
  • 19
    Personal Blog

    Personal Blog

    One article per week, the content is concise, neither salty nor light

    Personal Blog holds the source structure and article index for the author’s personal technical blog, which is closely tied to the “芋道源码” WeChat public account. It uses Markdown files and a static-site setup (with configuration like _config.yml) to organize posts about Java back-end engineering, distributed systems, and source-code deep dives. The README and index emphasize that the blog (in this repo) is paused and that new content is primarily delivered via the WeChat channel, but the GitHub repo remains the canonical archive of links and categories. ...
    Downloads: 0 This Week
    Last Update:
    See Project
  • 20
    libtdata

    libtdata

    Libtdata is a C library implements trees, index allocation and bit ops

    Libtdata is a small and portable C library implements a set of various search data structures, bit operations and index allocation.
    Downloads: 0 This Week
    Last Update:
    See Project
  • 21
    Invenio

    Invenio

    Invenio digital library framework

    Invenio is a highly customizable open-source framework for building large-scale digital repositories and research data platforms. Developed by CERN, it is designed to manage, index, and provide access to metadata-rich content such as publications, datasets, and multimedia files. Invenio provides a modular architecture, making it suitable for libraries, archives, and research institutions.
    Downloads: 0 This Week
    Last Update:
    See Project
  • 22
    pyFileSearcher

    pyFileSearcher

    simple searching tool for big fileservers

    pyFileSearcher was designed to be lightweight, easy to use, but capable of handling a large volume of files tool. A tool that I personally could use on large corporate servers to find out - which files have taken all my space in the last few days? It's free, it's opensource, it's for linux and windows. The program is written in Python 3 using the Qt5.
    Downloads: 0 This Week
    Last Update:
    See Project
  • 23
    Python Data Science Tutorials

    Python Data Science Tutorials

    Common data analysis and machine learning tasks using python

    ...Additional material addresses text mining, sentiment analysis, serialization with pickle, AutoML, regular expressions, and web scraping. The repository is organized by topic so learners can use it as a study roadmap or troubleshooting index. Many links document the earlier Python data science ecosystem, making the project especially valuable as a broad historical resource collection.
    Downloads: 0 This Week
    Last Update:
    See Project
  • 24
    Spatial Viewer for Oracle SQL Developer
    The purpose of GeoRaptor project is to extend Oracle SQL Developer with additional functionality for database administrators or developers working with Oracle Spatial data. Dear Users, A port to version 4.x of SQL Developer is being attempted. It is not simple due to the lack of development resources. Please keep using GeoRaptor with SQL Developer 3.x until the new version is available. Thanks. Simon Greener
    Downloads: 0 This Week
    Last Update:
    See Project
  • 25
    VividSTORM

    VividSTORM

    Correlated confocal and SMLM data visualization and analysis

    VividSTORM is a free and open-source standalone software with graphical user interface, for the correlated visualization and analysis of superresolution single molecule localization microscopy (SMLM) molecule lists and conventional pixel intensity-based images. The localization points (LPs) within this ROI can be analyzed using the selected built-in functions. NOTE: If you encounter issues not addressed by the user guide, please contact by message on this site or via e-mail for...
    Downloads: 0 This Week
    Last Update:
    See Project