Search Results for "cursive image text" - Page 6

Showing 249 open source projects for "cursive image text"

View related business solutions
  • Veeam Data Platform v13.1 - Get Your Free Trial Icon
    Veeam Data Platform v13.1 - Get Your Free Trial

    Secure by design, portable by default. Recover clean, fast, anywhere. Start a free trial.

    Try Veeam Data Platform today. Experience the unified platform that's secure by design, portable by default, and proven to recover clean, fast, and anywhere.
    Try it Free
  • MongoDB Atlas runs apps anywhere Icon
    MongoDB Atlas runs apps anywhere

    Deploy in 115+ regions with the modern database for every enterprise.

    MongoDB Atlas gives you the freedom to build and run modern applications anywhere—across AWS, Azure, and Google Cloud. With global availability in over 115 regions, Atlas lets you deploy close to your users, meet compliance needs, and scale with confidence across any geography.
    Start Free
  • 1
    Skywork-R1V4

    Skywork-R1V4

    Skywork-R1V is an advanced multimodal AI model series

    Skywork-R1V is an open-source multimodal reasoning model designed to extend the capabilities of large language models into vision-language tasks that require complex logical reasoning. The project introduces a model architecture that transfers the reasoning abilities of advanced text-based models into visual domains so the system can interpret images and perform multi-step reasoning about them. Instead of retraining both language and vision models from scratch, the framework uses a...
    Downloads: 0 This Week
    Last Update:
    See Project
  • 2
    ML Ferret

    ML Ferret

    Refer and Ground Anything Anywhere at Any Granularity

    Ferret is Apple’s end-to-end multimodal large language model designed specifically for flexible referring and grounding: it can understand references of any granularity (boxes, points, free-form regions) and then ground open-vocabulary descriptions back onto the image. The core idea is a hybrid region representation that mixes discrete coordinates with continuous visual features, so the model can fluidly handle “any-form” referring while maintaining precise spatial localization. The repo presents the vision-language pipeline, model assets, and paper resources that show how Ferret answers questions, follows instructions, and returns grounded outputs rather than just text.
    Downloads: 0 This Week
    Last Update:
    See Project
  • 3
    Large Concept Model

    Large Concept Model

    Language modeling in a sentence representation space

    ...It organizes training around concepts (rather than just raw labels), encouraging models to understand attributes, relations, and compositional structure that transfer across tasks. The repository provides training loops, data tooling, and evaluation routines to learn and probe these concept embeddings, typically from large image–text or weakly supervised corpora. It includes utilities to build concept vocabularies, map supervision signals to those vocabularies, and measure zero-shot or few-shot generalization. Probing tools help diagnose what the model knows—e.g., attribute recognition, relation understanding, or compositionality—so you can iterate on data and objectives. ...
    Downloads: 0 This Week
    Last Update:
    See Project
  • 4
    Stable Diffusion web UI for AMDGPUs

    Stable Diffusion web UI for AMDGPUs

    Stable Diffusion WebUI optimized for AMD GPUs with editing tools

    Stable Diffusion WebUI AMDGPU is a browser-based interface for generating images using Stable Diffusion, built with Gradio and adapted for AMD graphics hardware. It provides both text-to-image and image-to-image workflows, allowing users to create, refine, and upscale visuals within a single interface. It includes tools such as inpainting and outpainting for editing specific areas of an image, along with features like prompt matrix generation and attention controls to fine-tune outputs. Users can emphasize or de-emphasize elements in prompts to influence results more precisely. ...
    Downloads: 13 This Week
    Last Update:
    See Project
  • Build Agents and Models on One Platform Icon
    Build Agents and Models on One Platform

    Everything you need to build production-ready agents and models. Access 200+ Google and third-party AI models and tools.

    Gemini Enterprise Agent Platform is Google Cloud's comprehensive platform for developers to build, scale, govern, and optimize agents and models. Choose from Google's most advanced models and third-party models like Anthropic's Claude Model Family.
    Start Free
  • 5
    Stable Diffusion

    Stable Diffusion

    High-Resolution Image Synthesis with Latent Diffusion Models

    Stable Diffusion Version 2. The Stable Diffusion project, developed by Stability AI, is a cutting-edge image synthesis model that utilizes latent diffusion techniques for high-resolution image generation. It offers an advanced method of generating images based on text input, making it highly flexible for various creative applications. The repository contains pretrained models, various checkpoints, and tools to facilitate image generation tasks, such as fine-tuning and modifying the models. ...
    Leader badge
    Downloads: 225 This Week
    Last Update:
    See Project
  • 6
    File Sorter for Photographers

    File Sorter for Photographers

    Organize files/images from a csv or xlsx file.

    A user-friendly application to efficiently sort all types of files from a source folder into a destination folder based on a list of filenames provided in an Excel or CSV file.
    Downloads: 4 This Week
    Last Update:
    See Project
  • 7
    Phenaki - Pytorch

    Phenaki - Pytorch

    Implementation of Phenaki Video, which uses Mask GIT

    ...This repository will also endeavor to allow the researcher to train on text-to-image and then text-to-video. Similarly, for unconditional training, the researcher should be able to first train on images and then fine tune on video.
    Downloads: 4 This Week
    Last Update:
    See Project
  • 8
    AnimateDiff

    AnimateDiff

    Plug-n-play module turning text-to-image models into animation

    AnimateDiff is an open-source project designed to enhance text-to-image diffusion models by adding animation capabilities. It allows users to turn static images generated by popular text-to-image models into animated sequences without requiring additional model training. This plug-and-play tool is compatible with a wide range of community models and facilitates the generation of animation directly from pre-existing text-to-image models. ...
    Downloads: 24 This Week
    Last Update:
    See Project
  • 9
    KoboldCpp

    KoboldCpp

    Run GGUF models easily with a UI or API. One File. Zero Install.

    KoboldCpp is an easy-to-use AI text-generation software for GGML and GGUF models, inspired by the original KoboldAI. It's a single self-contained distributable that builds off llama.cpp and adds many additional powerful features.
    Leader badge
    Downloads: 944 This Week
    Last Update:
    See Project
  • Cut Data Warehouse Costs by 54% Icon
    Cut Data Warehouse Costs by 54%

    Easily migrate from Snowflake, Redshift, or Databricks with free tools.

    BigQuery delivers 54% lower TCO with exabyte scale and flexible pricing. Free migration tools handle the SQL translation automatically.
    Start Free
  • 10
    Mice MX OS speech to text Voice Control

    Mice MX OS speech to text Voice Control

    Mice speech to text with MX Cinnamon OS ISO

    Note about this image This image contains a system based on Linux MX, which was created to improve accessibility within the Linux environment. The distribution uses the Cinnamon desktop interface, which is configured to be operated using voice commands and outputs. The user interface and the control of your own devices and home automation systems can be customized and extended. The voice control program MiceStTM.py was developed to enable easy adaptation to other languages. However, only...
    Downloads: 10 This Week
    Last Update:
    See Project
  • 11

    realwatermark

    A Python application to add watermarks (text or image) to PDF files

    A Python application to add watermarks (text or image) to PDF files, converts them into image and back to PDF with options for OCR and compression.
    Downloads: 0 This Week
    Last Update:
    See Project
  • 12
    InternGPT

    InternGPT

    Open source demo platform where you can easily showcase your AI models

    InternGPT is an open-source multimodal AI framework designed to extend large language models beyond text interactions into visual reasoning and image manipulation tasks. The system integrates conversational AI with computer vision models so users can interact with images, videos, and visual environments through natural language instructions. Unlike traditional chat systems that rely solely on text prompts, InternGPT allows users to interact with visual content using both language and nonverbal signals such as pointing or highlighting objects within images. ...
    Downloads: 0 This Week
    Last Update:
    See Project
  • 13
    LlamaGen

    LlamaGen

    Autoregressive Model Beats Diffusion

    ...The project explores how scaling autoregressive models and improving image tokenization techniques can produce competitive results compared with modern diffusion-based image generators. LlamaGen provides several pre-trained models and training configurations that support both class-conditional image generation and text-conditioned image synthesis. The repository includes image tokenizers, training scripts, and models ranging from hundreds of millions to several billion parameters.
    Downloads: 1 This Week
    Last Update:
    See Project
  • 14
    Free AI Watermark Remover - FreeRepair

    Free AI Watermark Remover - FreeRepair

    AI-powered tool to quickly remove watermarks from images flawlessly

    AI Watermark Remover (Free And Open-Source) & Make Blurry Images Clearer Or Larger Tool - FreeRepair, Simulation IOPaint Based On The Django Of Python With No Sign-Up. As a free, open-source, AI-powered tool, FreeRepair makes it easy to remove watermarks, logos, text or clutter from images, and blurry images can be made clearer or larger. No installation, no internet connection, it works out of the box, safe and secure, unlimited.
    Leader badge
    Downloads: 68 This Week
    Last Update:
    See Project
  • 15
    VeilClip

    VeilClip

    Offline clipboard manager for Windows with history, search, and locked

    VeilClip is an open-source, offline clipboard manager for Windows 10 and Windows 11. It stores copied text, links, images, and file paths locally on your PC so you can search, pin, edit, reuse, and protect them later without a cloud account. Main features: - Clipboard history for text, links, images, and file paths - Fast search by content and source application - Favorites and pinned reusable items - Locked Notes protected by a PIN - Built-in text and image editing - Local backup, export, and import - Multi-language interface - Offline-first and privacy-focused design VeilClip uses local storage and does not require a cloud account for normal use.
    Downloads: 0 This Week
    Last Update:
    See Project
  • 16
    Pdf_tools
    ✅ Image to PDF Convert multiple image files into a single PDF. Supports formats: JPG, JPEG, PNG, BMP, TIFF. ✅ PDF Merger Merge multiple PDF files into one. Reorder PDF files before merging. ✅ PDF Splitter Split PDF files by range or into individual pages. ✅ Page Remover Remove specific pages from a PDF. ✅ Fill & Sign Add text and signature to a PDF.
    Downloads: 6 This Week
    Last Update:
    See Project
  • 17
    PII-Blackout

    PII-Blackout

    100% offline, AI-powered PDF redaction

    ...PII Blackout automatically scans, detects, and blackouts sensitive data points across your documents in one click. Absolute, Irreversible Security (Image-Level Blackout) Unlike standard PDF editors that merely place a black shape over editable text (which can easily be copied or uncovered), PII Blackout flattens and bakes the redaction directly into the image surface of the document. The covered data is permanently destroyed and mathematically impossible to recover. 100% Offline & Local Processing
    Downloads: 12 This Week
    Last Update:
    See Project
  • 18
    xSTUDIO

    xSTUDIO

    xSTUDIO is a high performance playback and review tool.

    xSTUDIO is a high performance playback and review tool designed by and for Visual Effects, Animation and Post Production professionals. The application can load and play large collections of media files. The efficient playback engine allows you to quickly load and play high resolution image formats with a wide range of file formats and encoding. Intuitive tools allow you to create and organise playlists and media sub-sets within playlists to build interactive review sessions, image and video...
    Downloads: 39 This Week
    Last Update:
    See Project
  • 19
    VideoCrafter2

    VideoCrafter2

    Overcoming Data Limitations for High-Quality Video Diffusion Models

    VideoCrafter is an open-source video generation and editing toolbox designed to create high-quality video content. It features models for both text-to-video and image-to-video generation. The system is optimized for generating videos from textual descriptions or still images, leveraging advanced diffusion models. VideoCrafter2, an upgraded version, improves on its predecessor by enhancing motion dynamics and concept combinations, especially in low-data scenarios. Users can explore a wide range of creative possibilities, producing cinematic videos that combine artistic styles and real-world scenes.
    Downloads: 6 This Week
    Last Update:
    See Project
  • 20
    Protect my docs

    Protect my docs

    Protect My Docs is a free watermarking tool for PDF and i

    Protect My Docs is a free watermarking tool for PDF and image files. Add a fully customizable text watermark with AES-128 encryption, ZIP AES-256 export, and mail optimization. No Python installation required. Available for Windows, macOS, and Linux.
    Downloads: 3 This Week
    Last Update:
    See Project
  • 21
    Gem Measure

    Gem Measure

    Live gemstone symmetry & distortion measurement

    Gem Measure overlays an adjustable polygon onto a gemstone image or video feed and instantly calculates edge length deviations, symmetry, and overall quality. It's designed for jewelers, gemologists, and hobbyists who need quick, accurate measurements without expensive software. Works with USB microscopes, webcams, or any image file. How to Use Open an image (F) or connect a camera (C).
    Downloads: 1 This Week
    Last Update:
    See Project
  • 22
    KherveTeX

    KherveTeX

    Visual LaTeX editor: edit like Word, output like LaTeX, track with Git

    KherveTeX is a document editor that produces real LaTeX and tracks every change with Git — giving you the ease of a word processor and the typesetting quality of LaTeX, without ever touching a .tex file by hand. Write documents visually with familiar formatting tools. KherveTeX maintains a clean document model and serializes it to LaTeX, then compiles to PDF via the bundled Tectonic engine — no external LaTeX installation required. The PDF preview updates as you work. Every save is an...
    Downloads: 4 This Week
    Last Update:
    See Project
  • 23
    Folio - macOS Ebook & Document Reader

    Folio - macOS Ebook & Document Reader

    Offline ebook reader for macOS: PDF, DjVu, EPUB, FB2, MOBI & more.

    Folio is a free, open-source, offline ebook and document reader for macOS Apple Silicon. Read PDF, DjVu, EPUB, FB2, MOBI, PRC, XPS, OXPS, CBZ, TXT and image files in one native desktop application. Folio includes continuous vertical scrolling, fast lazy page rendering, bookmarks, text search, zoom controls, fit-to-width and fit-to-height modes, print preview, reading-position history and optimized DjVu-to-PDF export. Books stay on your Mac — no account, cloud upload or telemetry is required. ...
    Downloads: 2 This Week
    Last Update:
    See Project
  • 24
    surtr automation v4.0

    surtr automation v4.0

    Visual desktop automation, OCR, JSON & web automation toolkit

    Surtr is a free, open-source Windows automation platform that combines computer vision, OCR, scripting, web automation, and developer tools into one powerful toolkit. Automate applications using image recognition instead of fragile screen coordinates, extract text from screenshots, receipts, and invoices with OCR, record actions into reusable scripts, build workflows visually in SurtrUI, or control everything remotely through the built-in WebUI. Surtr also includes a high-performance downloader and web scraper, native JSON creation and parsing, REST API integration, task scheduling, file automation, and a powerful CLI scripting engine with variables, loops, and conditions. ...
    Downloads: 1 This Week
    Last Update:
    See Project
  • 25
    dashAI

    dashAI

    dashAI: an interactive platform for training, evaluating and deploying

    dashAI is an open-source, No-code workbench for Exploratory Data Analysis and classical ML. Visual data preparation, multi-model experiments, XAI explainability, and a plugin-based extensible catalog. The platform guides users through a complete, traceable workflow — data ingestion → visual exploration → preprocessing → model training → evaluation → explainability — without writing a single line of code. Each step is explicit and reversible, keeping the user in control rather than...
    Leader badge
    Downloads: 8 This Week
    Last Update:
    See Project