Search Results for "text code" - Page 3

Sort By:

Showing 343 open source projects for "text code"

View related business solutions

Python Clear Filters & Widen Search

$300 Free Credits for Your Google Cloud Projects
Start building on Google Cloud with $300 in free credits. No commitment, no credit card required until you're ready to scale.

Launch your next project with $300 in free Google Cloud credits—no strings attached. Test, build, and deploy without risk. Use your credits across the entire Google Cloud platform to find what works best for your needs. After your credits are used, continue with always-free tier services. Only pay when you're ready to scale. Sign up in minutes and start exploring.

Start Free Trial
Our Free Plans just got better! | Auth0
With up to 25k MAUs and unlimited Okta connections, our Free Plan lets you focus on what you do best—building great apps.

You asked, we delivered! Auth0 is excited to expand our Free and Paid plans to include more options so you can focus on building, deploying, and scaling applications without having to worry about your security. Auth0 now, thank yourself later.

Try free now
1

Unsloth Studio

Unified web UI for training and running open models locally

...Users can interact with models through chat, upload files like PDFs or images, and execute code within the environment to improve outputs. By combining powerful optimization techniques with an intuitive UI, Unsloth Studio simplifies the process of building and customizing AI models locally.

Downloads: 11 This Week

Last Update: 4 days ago
See Project
2

HiDream-I1

Open-source image generative foundation model

HiDream-I1 is an open-source image generation foundation model with 17 billion parameters. It is designed to produce high-quality images from text prompts while keeping inference practical through efficient model design. The project provides full, dev, and fast model variants with different inference step counts. It supports direct Python inference scripts, an interactive Gradio demo, and integration through the Hugging Face Diffusers library. The model uses a Llama 3.1 text encoder path and...

Downloads: 2 This Week

Last Update: 2026-06-17
See Project
3

NetworkX

Network analysis in Python

...Many standard graph algorithms. Network structure and analysis measures. Generators for classic graphs, random graphs, and synthetic networks. Nodes can be "anything" (e.g., text, images, XML records). Edges can hold arbitrary data (e.g., weights, time-series). Open source 3-clause BSD license. Well tested with over 90% code coverage. Additional benefits from Python include fast prototyping, easy to teach, and multi-platform. Find the shortest path between two nodes in an undirected graph. Python’s None object is not allowed to be used as a node. ...

Downloads: 5 This Week

Last Update: 2025-12-08
See Project
4

rich

Rich is a Python library for rich text and beautiful formatting

The Rich API makes it easy to add color and style to terminal output. Rich can also render pretty tables, progress bars, markdown, syntax highlighted source code, tracebacks, and more, out of the box. Rich is a Python library for rich text and beautiful formatting in the terminal. Rich works with Linux, OSX, and Windows. True color/emoji works with new Windows Terminal, classic terminal is limited to 16 colors. Rich requires Python 3.7 or later. Effortlessly add rich output to your application, you can import the rich print method, which has the same signature as the builtin Python function. ...

Downloads: 3 This Week

Last Update: 2026-04-12
See Project
AI-powered service management for IT and enterprise teams
Enterprise-grade ITSM, for every business

Give your IT, operations, and business teams the ability to deliver exceptional services—without the complexity. Maximize operational efficiency with refreshingly simple, AI-powered Freshservice.

Try it Free
5

GLM-OCR

Accurate × Fast × Comprehensive

GLM-OCR is an open-source multimodal optical character recognition (OCR) model built on a GLM-V encoder–decoder foundation that brings robust, accurate document understanding to complex real-world layouts and modalities. Designed to handle text recognition, table parsing, formula extraction, and general information retrieval from documents containing mixed content, GLM-OCR excels across major benchmarks while remaining highly efficient with a relatively compact parameter size (~0.9B), enabling deployment in high-concurrency services and edge environments. The model’s multimodal capabilities allow it to reason across image and text content holistically, capturing structured and unstructured information from pages that include dense tables, seals, code snippets, and varied document graphics. ...

Downloads: 8 This Week

Last Update: 2026-04-08
See Project
6

DeerFlow

Deep Research framework, combining language models with tools

DeerFlow is an open-source, community-driven “deep research” framework / multi-agent orchestration platform developed by ByteDance. It aims to combine the reasoning power of large language models (LLMs) with automated tool-use — such as web search, web crawling, Python execution, and data processing — to enable complex, end-to-end research workflows. Instead of a monolithic AI assistant, DeerFlow defines multiple specialized agents (e.g. “planner,” “searcher,” “coder,” “report generator”)...

Downloads: 39 This Week

Last Update: 2026-06-25
See Project
7

PersonaPlex

PersonaPlex code

PersonaPlex is an open-source real-time conversational speech AI model that goes beyond traditional text chat by providing full-duplex speech-to-speech interaction, meaning it can listen and talk at the same time instead of waiting for you to finish speaking before responding. This architectural approach eliminates awkward pauses and makes conversations feel much more human-like, with natural behaviors such as overlapping speech, interruptions, and fluent turn-taking, traits that traditional...

Downloads: 0 This Week

Last Update: 2026-03-02
See Project
8

Tiktoken

tiktoken is a fast BPE tokeniser for use with OpenAI's models

tiktoken is a high-performance, tokenizer library (based on byte-pair encoding, BPE) designed for use with OpenAI’s models. It handles encoding and decoding text to token IDs efficiently, with minimal overhead. Because tokenization is a fundamental step in preparing text for models, tiktoken is optimized for speed, memory, and correctness in model contexts (e.g. matching OpenAI’s internal tokenization). The repo supports multiple encodings (e.g. “cl100k_base”) and lets users switch encoding names to match different model contexts. ...

Downloads: 6 This Week

Last Update: 2026-05-15
See Project
9

Think Bayes 2

Text and code for the second edition of Think Bayes, by Allen Downey

...It teaches Bayesian reasoning through computational methods instead of relying mainly on symbolic mathematics. Each chapter is presented as a Jupyter notebook where readers can study the text, run examples, and complete exercises. Separate solution materials help learners check their work and explore alternative approaches. The lessons cover probability distributions, Bayesian updating, estimation, prediction, comparison, and decision-making. Notebooks can run in Google Colab or be downloaded for local use. The repository also contains book sources, supporting code, and environment files for reproducible study.

Downloads: 1 This Week

Last Update: 2 days ago
See Project
Earn up to 16% annual interest with Nexo.
Access competitive interest rates on your digital assets.

Generate interest, borrow against your crypto, and trade a range of cryptocurrencies — all in one platform. Geographic restrictions, eligibility, and terms apply.

Get started with Nexo.
10

Papermerge

Open Source Document Management System for Digital Archives

...OCR technology is vital part of Papermerge. It extracts text information from scanned documents, PDF, JPEG, TIFF files.

Downloads: 21 This Week

Last Update: 2025-07-24
See Project
11

TextFSM

Python module for parsing semi-structured text into python tables

TextFSM is a Python library created by Google that provides a template-based state machine engine for parsing semi-structured text. It is particularly useful for extracting structured data from command-line interface (CLI) outputs, such as those from network devices, routers, and switches. By defining parsing logic through reusable template files, TextFSM transforms unstructured text into structured data like lists or tables without requiring complex regular expression code. ...

Downloads: 0 This Week

Last Update: 2025-10-11
See Project
12

AutoGluon

AutoGluon: AutoML for Image, Text, and Tabular Data

AutoGluon enables easy-to-use and easy-to-extend AutoML with a focus on automated stack ensembling, deep learning, and real-world applications spanning image, text, and tabular data. Intended for both ML beginners and experts, AutoGluon enables you to quickly prototype deep learning and classical ML solutions for your raw data with a few lines of code. Automatically utilize state-of-the-art techniques (where appropriate) without expert knowledge. Leverage automatic hyperparameter tuning, model selection/ensembling, architecture search, and data processing. ...

Downloads: 6 This Week

Last Update: 2025-12-19
See Project
13

Remarkable for Linux

The Markdown Editor for Linux

With Live Preview you can see your changes as you make them. There is no need to export first to check your syntax. This is accompanied by synchronized scrolling. Remarkable has Github Flavoured Markdown. This has a simple, easy-to-learn syntax with features like checklists, highlighting, links, images and more. Remarkable allows you to export your files to PDF and HTML from within the app. The HTML code is even prettified and PDFs have a TOC. You can style your markdown documents however...

Downloads: 0 This Week

Last Update: 2024-09-22
See Project
14

The Hundred-Page Machine Learning Book

The Python code to reproduce illustrations from Machine Learning Book

The Hundred-Page Machine Learning Book is the official companion repository for The Hundred-Page Machine Learning Book written by machine learning researcher Andriy Burkov. The repository contains Python code used to generate the figures, visualizations, and illustrative examples presented in the book. Its purpose is to help readers better understand the concepts explained in the text by allowing them to run and experiment with the underlying code themselves. The book itself provides a concise overview of machine learning theory and practice, covering topics such as supervised learning, unsupervised learning, neural networks, and optimization algorithms. ...

Downloads: 3 This Week

Last Update: 2026-03-12
See Project
15

HunyuanOCR

OCR expert VLM powered by Hunyuan's native multimodal architecture

HunyuanOCR is an open-source, end-to-end OCR (optical character recognition) Vision-Language Model (VLM) developed by Tencent‑Hunyuan. It’s designed to unify the entire OCR pipeline, detection, recognition, layout parsing, information extraction, translation, and even subtitle or structured output generation, into a single model inference instead of a cascade of separate tools. Despite being fairly lightweight (about 1 billion parameters), it delivers state-of-the-art performance across a...

Downloads: 2 This Week

Last Update: 4 days ago
See Project
16

video-use

Edit videos with Claude Code

...By combining structured text representations with on-demand visual previews, it minimizes processing overhead while maintaining high-quality results. Overall, Video Use reimagines video editing as an AI-driven, conversational workflow.

Downloads: 32 This Week

Last Update: 2026-07-01
See Project
17

Dia

A TTS model capable of generating ultra-realistic dialogue

Dia is a neural text-to-speech model designed specifically for generating ultra-realistic dialogue in a single pass. Instead of focusing on isolated sentences or flat narration, it is optimized for conversational audio, complete with natural turn-taking, prosody, and pacing. The model can be conditioned on a reference audio sample, allowing you to control emotion, tone, and other stylistic aspects of the speech.

Downloads: 1 This Week

Last Update: 2025-11-28
See Project
18

FireRedTTS-2

Long-form streaming TTS system for multi-speaker dialogue generation

FireRedTTS2 is a next-generation open-source text-to-speech (TTS) system focused on long-form, streaming speech synthesis for multi-speaker dialogue, delivering stable natural speech with context-aware prosody and reliable speaker transitions that support real-time and conversational applications. It features a specialized streaming speech tokenizer and a dual-transformer architecture that enables low latency and high-quality synthesis, making it suitable for interactive systems like...

Downloads: 0 This Week

Last Update: 2026-02-16
See Project
19

MiniMind-V

"Big Model" trains a visual multimodal VLM with 26M parameters

MiniMind-V is an experimental open-source project that aims to train a very small multimodal vision–language model (VLM) from scratch with extremely low compute and cost, making research and experimentation accessible to more people. The repository showcases training workflows and code designed to produce a 26-million parameter model—including both image and text capabilities—using minimal resources in very little time, reflecting a trend toward democratizing AI research. MiniMind-V combines techniques from modern vision-language modeling but focuses on efficiency and simplicity so that individuals or small teams can explore multimodal learning without massive GPU clusters. ...

Downloads: 1 This Week

Last Update: 2026-01-21
See Project
20

o1-engineer

o1-engineer is a command-line tool designed to assist developers

o1-engineer is a command-line development assistant powered by OpenAI’s API. It helps developers interact with projects through commands for code generation, file editing, project planning, and code review. The tool can add, edit, and manage both files and folders directly from the terminal. Its planning command can create structured project plans that can then guide systematic file and directory generation. It also keeps conversation history and allows users to save or reset context as...

Downloads: 0 This Week

Last Update: 2026-06-24
See Project
21

LLM Foundry

LLM training code for MosaicML foundation models

Introducing MPT-7B, the first entry in our MosaicML Foundation Series. MPT-7B is a transformer trained from scratch on 1T tokens of text and code. It is open source, available for commercial use, and matches the quality of LLaMA-7B. MPT-7B was trained on the MosaicML platform in 9.5 days with zero human intervention at a cost of ~$200k. Large language models (LLMs) are changing the world, but for those outside well-resourced industry labs, it can be extremely difficult to train and deploy these models. ...

Downloads: 9 This Week

Last Update: 2025-07-29
See Project
22

amrlib

A python library that makes AMR parsing, generation and visualization

A python library that makes AMR parsing, generation and visualization simple. amrlib is a python module designed to make processing for Abstract Meaning Representation (AMR) simple by providing the following functions. Sentence to Graph (StoG) parsing to create AMR graphs from English sentences. Graph to Sentence (GtoS) generation for turning AMR graphs into English sentences. A QT-based GUI to facilitate the conversion of sentences to graphs and back to sentences. Methods to plot AMR graphs...

Downloads: 0 This Week

Last Update: 2026-03-07
See Project
23

ThinkStats2

Text and supporting code for Think Stats, 2nd Edition

ThinkStats2 is the code and text companion for the second edition of Think Stats, an introduction to statistics and data science for Python programmers. It teaches probability and statistical reasoning through short programs, experiments, and analysis of real datasets. The material emphasizes exploratory methods that help readers ask and answer practical questions with data.

Downloads: 0 This Week

Last Update: 2 days ago
See Project
24

MiniMind-O

A 0.1B Omni model trained from scratch

MiniMind-O is an educational open-source project for building a small end-to-end Omni model from scratch. It extends the MiniMind family by exploring a model that can handle text, audio, and image inputs while producing text and streaming speech outputs. The project is designed to make multimodal AI training more accessible by keeping the model size small enough for ordinary personal hardware. It includes both mini and full training data paths, allowing learners to run a complete workflow quickly or reproduce the released model setup more closely. ...

Downloads: 0 This Week

Last Update: 2026-06-28
See Project
25

Upscale-A-Video

Temporal-Consistent Diffusion Model for Real-World Video

Upscale-A-Video is a diffusion-based video super-resolution project from the CVPR 2024 Highlight paper “Temporal-Consistent Diffusion Model for Real-World Video Super-Resolution.” It upscales low-resolution videos while using text prompts to guide the enhancement process. The model is designed for real-world videos where compression artifacts, blur, aging, or generated-video defects can make ordinary upscaling less reliable. The repository includes inference code, example inputs, configuration structure, pretrained model instructions, and optional LLaVA-assisted prompt support. ...

Downloads: 4 This Week

Last Update: 2026-07-02
See Project