Deploy in 115+ regions with the modern database for every enterprise.
MongoDB Atlas gives you the freedom to build and run modern applications anywhere—across AWS, Azure, and Google Cloud. With global availability in over 115 regions, Atlas lets you deploy close to your users, meet compliance needs, and scale with confidence across any geography.
Start Free
Ship Agents Faster
Transform your applications and workflows into powerful agentic systems at global scale.
Gemini Enterprise Agent Platform lets you rapidly build, scale, govern and optimize production-ready agents grounded in your organization's data. The platform enables developers to build custom or pre-built agents for virtually any use case. New customers get $300 in free credits.
Project B is a platform for various Bible programs using Java. It will support desktop applications like the On-Line Bible and Sword, there is a Servlet interface, some add-in macros for MS Word. Other interfaces are in development.
WebSPHINX is a web crawler (robot, spider) Javaclass library, originally developed by Robert Miller of Carnegie Mellon University. Multithreaded, tollerant HTML parsing, URL filtering and page classification, pattern matching, mirroring, and more.
This project is a Python-based HTTP web proxy server that hooks into MySQL to store a full history of your browsing. Allows you to check out statistics about your browsing habits. Creates a personal portal page, has search features, multi-user, filters.
Voambolana (pronouce VOO-BOO-LUH-NUH) is an on-line dictionary that converts foreign languages to a native language. Voambolana uses SAX parser and XSLT transformer. The tools used includes Ant, Xerces, Xalan (XNI) and Apache from the Apache Group.
META Spider is a neural network powered search engine which aim is to determine if a given raw information is relevant, using the experience of the active user of the hosting computer as a reference.
The JSearch Project wants to provide the internet with a Java based generic interface for search engines. It consists of a core interface, search engine adaptors, a sort/merge module and a JSP based GUI.
XQuench is an XML Query parser and engine. The aim is to provide programmers with an API that implements the specifications at http://www.w3.org/XML/Query At first, XQuench will be Java-only. Future versions will include C++, while keeping a similar API.
New customers can spin up VMs, build with AI, and query data at no cost.
Put your $300 in credit toward real workloads, then keep building with free monthly usage for 20+ products. No commitment and no charge until you upgrade.
Frosttie (FROnt-end SchemaTron Text Internet Engine) takes XHTML pages and processes them with various user-definable filters such a W3C's WAI, Section 508 (US) web usability compliance, ad removal, etc. It can be used with zKnowMan.
This project contains all the code for the eXploringXML column on WebReference.com at http://exploringxml.com .
Currently this is only an applet for parsing and displaying Rich Site Summary (RSS) files, but more Java code for XML will come
jMarks is a full-blown multi-user web-based bookmark solution, written in Java. jMarks allows people to mark their online bookmarks as public or private, and can track the last time each bookmarked site was updated.
Omseek has been renamed to Xapian. Xapian is a Search Engine Library, written in C++ with bindings for Perl, Python, PHP, Java, Tcl, C# and Ruby. It allows you to easily add advanced indexing and search facilities to your applications.
Competence will be an expandable information retrieval system.
Like the non-free Glimpse/WebGlimpse Competence will index various kinds of documents stored locally or on the web and provide an easy-to-use interface for search and retrieval.
The New Wave Searchables are a framework for implementing scalable search services.
They will allow searching deep web contents, implement best practices from the field
of information retrieval (formal query capability descriptions, query transformation
MediaHunter consists of a server/db software and a frontend plug-in for java gnutella servents. The server works like a cddb mirror and uses XML protocol for lookups, the plugin queries the server and manages searches and downloads on the gnutella servent
This is a simple java based interface to the Open Directory Project. (www.dmoz.org) The javaclass supplied can retrieve data from dmoz on a request per request basis to give your site access to dmoz data.
JoBo is a web site mirroring tool. It has a graphical UI but there is a also command line version. Supports robot exclusion protocol (but this can be disabled)
NeatSeeker is a collection of Java classes for building search
engines. It also offers reference implementations of an HTML indexer
and a Servlet API 2.2 compliant Java servlet that can be used for
indexing and searching web sites.
JSherlock is a clone of the well known application "Sherlock" on MacOS.
It is a frontend for querying various internet-search websites such as yahoo, google, freshmeat .... JSherlock is extendable via plugins.
PyEsp - Enhanced/Evolving/Extensible Semantic Profiling.
This Python program will sort and filter search results by applying semantic profiling on web pages. The program will learn the user preferences and profiling will be done on the client computer.
Helping to bridge the gap between AJAX and Search Engine Optimization (SEO) in Java, to have an automated way of serving web crawlers with static pages that is dynamically generated from AJAX.