Retriever is a simple crawler packed as a Java library that allows developers to collect and manipulate documents reachable by a variety of protocols (e.g. http, smb). You'll easily crawl documents shared in a LAN, on the Web, and many other sources.

Project Activity

See All Activity >

Categories

Search Engines

License

Apache License V2.0

Follow Retriever: a light, extensible crawler

Retriever: a light, extensible crawler Web Site

Other Useful Business Software
Ship AI Apps Faster with Vertex AI Icon
Ship AI Apps Faster with Vertex AI

Go from idea to deployed AI app without managing infrastructure. Vertex AI offers one platform for the entire AI development lifecycle.

Ship AI apps and features faster with Vertex AI—your end-to-end AI platform. Access Gemini 3 and 200+ foundation models, fine-tune for your needs, and deploy with enterprise-grade MLOps. Build chatbots, agents, or custom models. New customers get $300 in free credit.
Try Vertex AI Free
Rate This Project
Login To Rate This Project

User Reviews

Be the first to post a review of Retriever: a light, extensible crawler!

Additional Project Details

Languages

English

Intended Audience

Developers

User Interface

Other toolkit

Programming Language

Java

Related Categories

Java Search Engines

Registered

2007-12-03