Search Results for "pdf data mining" - Page 12

Showing 311 open source projects for "pdf data mining"

View related business solutions
  • $300 Free Credits for Your Google Cloud Projects Icon
    $300 Free Credits for Your Google Cloud Projects

    Start building on Google Cloud with $300 in free credits. No commitment, no credit card required until you're ready to scale.

    Launch your next project with $300 in free Google Cloud credits—no strings attached. Test, build, and deploy without risk. Use your credits across the entire Google Cloud platform to find what works best for your needs. After your credits are used, continue with always-free tier services. Only pay when you're ready to scale. Sign up in minutes and start exploring.
    Start Free Trial
  • Build Securely on AWS with Proven Frameworks Icon
    Build Securely on AWS with Proven Frameworks

    Lay a foundation for success with Tested Reference Architectures developed by Fortinet’s experts. Learn more in this white paper.

    Moving to the cloud brings new challenges. How can you manage a larger attack surface while ensuring great network performance? Turn to Fortinet’s Tested Reference Architectures, blueprints for designing and securing cloud environments built by cybersecurity experts. Learn more and explore use cases in this white paper.
    Download Now
  • 1
    The Nheengatu Project is a Java library that provides HTML markup abstraction allowing you to reutilize it to generate PDF files, OpenOffice documents, image files, etc. The goal of this project is to maximize the use of HTML markup procedures.
    Downloads: 0 This Week
    Last Update:
    See Project
  • 2
    Connla is a Java library for creating data collections which can be exported to TXT, CSV, HTML, XHTML, XML, PDF and XLS formats.
    Downloads: 0 This Week
    Last Update:
    See Project
  • 3
    ESTminer is a Web application and database schema for interactive mining of expressed sequence tag (EST) contig and cluster data sets.
    Downloads: 0 This Week
    Last Update:
    See Project
  • 4
    Yawda is an MVC Web development application framework based on Struts. With it, you can easily output HTML,SVG,PNG,JPEG,RTF and PDF with data from several sources. It uses cayenne,rhino,itext,Batik,Struts,Velocity,jakarta commons.
    Downloads: 0 This Week
    Last Update:
    See Project
  • Our Free Plans just got better! | Auth0 Icon
    Our Free Plans just got better! | Auth0

    With up to 25k MAUs and unlimited Okta connections, our Free Plan lets you focus on what you do best—building great apps.

    You asked, we delivered! Auth0 is excited to expand our Free and Paid plans to include more options so you can focus on building, deploying, and scaling applications without having to worry about your security. Auth0 now, thank yourself later.
    Try free now
  • 5
    A collection of tools for Data Mining that aims to be Data Source independent
    Downloads: 0 This Week
    Last Update:
    See Project
  • 6
    ReportGUI is a Tool to design PDF report graphically. It uses FOP from Apache and supports data access, recursive SubDetails, bar codes, and other format properties. It too generate java source files to call from your code and show designed reports.
    Downloads: 0 This Week
    Last Update:
    See Project
  • 7
    txtkit is a visual text mining tool for exploring large amounts of multilingual texts. It's an multiuser-application which mainly focuses on the process of reading and reasoning as series of decisions and events.
    Downloads: 0 This Week
    Last Update:
    See Project
  • 8
    webExtractor is a Java application that is used for extracting specific content from web based HTML, XML, CSV, and free form text. The extracted data can be used for data gathering and mining purposes.
    Downloads: 3 This Week
    Last Update:
    See Project
  • 9
    A utility to extract data from RDBMSs and convert into .arff file format required by WEKA data mining tool set, both interactive wizard and batch working modes.
    Downloads: 0 This Week
    Last Update:
    See Project
  • Ship Agents Faster Icon
    Ship Agents Faster

    Transform your applications and workflows into powerful agentic systems at global scale.

    Gemini Enterprise Agent Platform lets you rapidly build, scale, govern and optimize production-ready agents grounded in your organization's data. The platform enables developers to build custom or pre-built agents for virtually any use case. New customers get $300 in free credits.
    Get Started Free
  • 10
    TagPrint is a DOM serialization library. It prints DOM documents with various format, such as XML, HTML, PDF, RTF, etc... You can write these documents very easily.
    Downloads: 0 This Week
    Last Update:
    See Project
  • 11
    This is a simple application that is meant to turn pdf files into xml files with a slant toward data management rather than visual appearance. This means that it is more suited toward data extraction than exact representation of data from a visual stand
    Downloads: 0 This Week
    Last Update:
    See Project
  • 12
    Weka-Parallel is a modification to Weka, created with the intention of being able to harness the power of Weka and the speed of parallel processing to be able to run a number of data mining and machine learning algorithms quickly.
    Downloads: 0 This Week
    Last Update:
    See Project
  • 13
    Powerful cataloguing software for various types of files (audio, video, various text documents, software packages etc.) based on XML technologies and thus providing broad capabilities for data manipulation and reporting (text, HTML/XHTML/PDF, RTF, whateve
    Downloads: 0 This Week
    Last Update:
    See Project
  • 14
    LearnML is a XML based markup language to put learning materials in the web. Based on a simple syntax, LearnML documents can be transformed to any kind of web page (HTML, XHTML) or (printable) PDF document.
    Downloads: 0 This Week
    Last Update:
    See Project
  • 15
    IDEA is a package for input and output of data out of/ into a database. Beginning as a web-application, IDEA generates your HTML-forms for the input and gives you some HTML- or PDF-output back. Everything IDEA does comes from one XML-file per form.
    Downloads: 0 This Week
    Last Update:
    See Project
  • 16
    Harvestman is a context aware metasearch engine which functions as a universal infromation gatherer and data mining system for the internet.
    Downloads: 0 This Week
    Last Update:
    See Project
  • 17
    Generates any PDF/TXT report without any headaches with a new geneartion report rendering engine. Merging two XML files (report layout XML, data XML) by this tool to give you any reports you want.
    Downloads: 0 This Week
    Last Update:
    See Project
  • 18
    Java doclet capable of generating Java source documentation in XML format. XSL stylesheets can be used to transform this XML data to other formats (e.g. HTML, PDF).
    Downloads: 0 This Week
    Last Update:
    See Project
  • 19
    Data Mining Platform is a platform for data mining and analysis. It contains many of the new and sophisticated methods such as kernel-based classification, two-way clustering, bayesian networks, pattern recognition for time series analysis and many other
    Downloads: 0 This Week
    Last Update:
    See Project
  • 20
    OpenGMP is an open service platform for implementing advanced decision support solutions for the mining enterprise.
    Downloads: 0 This Week
    Last Update:
    See Project
  • 21
    Eclipse PDF Renderer is a plugin for the Eclipse IDE. It adds a view to Eclipse in which PDF documents can be displayed. It might be useful if your Eclipse workspace contains several PDF files, or if you're using other Eclipse plugins like texlipse.
    Downloads: 0 This Week
    Last Update:
    See Project
  • 22
    A framework for creating incremental updates to PDF files, based on the excellent iText library. Additions and modifications to PDF files can be created and appended to existing PDFs, without re-writing the PDF file in its entirety.
    Downloads: 0 This Week
    Last Update:
    See Project
  • 23
    The iText based PDF sCramblEr allows you to encrypt a PDF using one or more public certificates of the addressees (one or more .cer files). For each .cer file, you can enforce specific PDF permissions: (dis)allow printing, (dis)allow modification,...
    Downloads: 0 This Week
    Last Update:
    See Project
  • 24
    Java based tool to convert HTML/DHTM to PDF document.
    Downloads: 0 This Week
    Last Update:
    See Project
  • 25
    This is a tool to convert pdf files to html/text files and extract images.
    Downloads: 0 This Week
    Last Update:
    See Project
Auth0 Logo