Showing 89 open source projects for "annotation"

View related business solutions
  • Build Agents and Models on One Platform Icon
    Build Agents and Models on One Platform

    Everything you need to build production-ready agents and models. Access 200+ Google and third-party AI models and tools.

    Gemini Enterprise Agent Platform is Google Cloud's comprehensive platform for developers to build, scale, govern, and optimize agents and models. Choose from Google's most advanced models and third-party models like Anthropic's Claude Model Family.
    Start Free
  • Build Data Resilience - Take the Assessment Today Icon
    Build Data Resilience - Take the Assessment Today

    Can you recover when it matters most? Take this quick assessment to identify gaps and build greater recovery confidence.

    Is your recovery strategy as strong as you think? Take this quick self-assessment to check your recovery readiness and gain tailored insights. In only 2 minutes, you'll learn where you fall on the recovery readiness scale.
    Take the Assessment
  • 1
    corona

    corona

    Reverse engineering SARS-CoV-2

    corona is an exploratory bioinformatics project that applies reverse-engineering ideas to SARS-CoV-2. It treats biological information processing as an analogy to software analysis to help explain the viral genome from first principles. Python scripts download genomic sequences from GenBank and translate RNA into amino-acid chains. The project identifies and annotates proteins encoded by the genome. It also experiments with OpenMM for molecular simulation and protein-folding work. Supporting...
    Downloads: 2 This Week
    Last Update:
    See Project
  • 2
    Fun4Me

    Fun4Me

    A package for functional annotation for metagenomes

    This package includes a few programs for rapid functional annotation for metagenomic sequences, including, 1) Gene prediction by FragGeneScan; 2) Similarity search by RAPSearch2; 3) Functional annotation in GO (Gene Ontology) and EC (Enzyme Commission) based on similarity search results; 4) From EC to metabolic pathway reconstruction by MinPath. Inputs: Just sequencing reads (or assemblies) Outputs: Protein-coding genes (or gene fragments); similarity search; functional annotations (in GO and EC); metabolic pathways.
    Downloads: 0 This Week
    Last Update:
    See Project
  • 3
    COCO Annotator

    COCO Annotator

    Web-based image segmentation tool for object detection & localization

    ...The annotation process is delivered through an intuitive and customizable interface and provides many tools for creating accurate datasets. Several annotation tools are currently available, with most applications as a desktop installation. Once installed, users can manually define regions in an image and creating a textual description. Generally, objects can be marked by a bounding box, either directly, through a masking tool, or by marking points to define the containing area. ...
    Downloads: 1 This Week
    Last Update:
    See Project
  • 4
    CBMPy

    CBMPy

    PySCeS Constraint Based Modelling

    PySCeS CBMPy is a new platform for constraint based modelling and analysis. It has been designed using principles developed in the PySCeS simulation software project: usability, flexibility and accessibility. CBMPy supports the latest standards for encoding CBM models encoding, SBML L3 FBC, COBRA as well as MIRIAM compliant RDF and custom annotations. Its architecture is both extensible and flexible using data structures that are intuitive to the biologist while transparently...
    Downloads: 2 This Week
    Last Update:
    See Project
  • PRTG Catches Network Issues Before They Cause Downtime Icon
    PRTG Catches Network Issues Before They Cause Downtime

    Threshold-based alerts flag problems early, so your team can act before users notice, not after.

    Reactive troubleshooting usually means hearing about a problem from frustrated users, not your monitoring tool. PRTG sets threshold-based alerts across devices, servers and applications, notifying your team by email, SMS or push the moment a metric crosses a set limit. That means catching a failing disk or overloaded server before it becomes an outage and getting time back from firefighting. Start a free trial and set your first alerts today.
    Download 30-Day Trial
  • 5
    TI2BioP allows mainly the calculation of topological indices (spectral moments) derived from inferred and artificial 2D structures of DNA, RNA and proteins being possible to carry out a structure-function correlation irrespective of sequence alignments. TI2BioP version 3.0 is a python platform with a graphical interface designed for Windows, Linux and Mac OS.
    Downloads: 0 This Week
    Last Update:
    See Project
  • 6
    LabelImg

    LabelImg

    Graphical image annotation tool and label object bounding boxes

    ...Virtualenv can avoid a lot of the QT / Python version issues. Build and launch using the instructions. Click 'Change default saved annotation folder' in Menu/File. Click 'Open Dir'. Click 'Create RectBox'. Click and release left mouse to select a region to annotate the rect box. You can use right mouse to drag the rect box to copy or move it. The annotation will be saved to the folder you specify. You can refer to the hotkeys to speed up your workflow.
    Downloads: 60 This Week
    Last Update:
    See Project
  • 7
    3D ResNets for Action Recognition

    3D ResNets for Action Recognition

    3D ResNets for Action Recognition (CVPR 2018)

    We uploaded the pretrained models described in this paper including ResNet-50 pretrained on the combined dataset with Kinetics-700 and Moments in Time. We significantly updated our scripts. If you want to use older versions to reproduce our CVPR2018 paper, you should use the scripts in the CVPR2018 branch.
    Downloads: 0 This Week
    Last Update:
    See Project
  • 8

    PPLine

    SNP calling, annotation and gene/transcripts expression quantification

    ...PPLine provides: - read mapping (STAR/Tophat2/bowtie/bowtie2), including novel splice junsctions discovery - gene and transcript expression estimation (HTSeq-count/Cufflinks) - SNP calling with BQSR and indel realignment (samtools/GATK) - variant annotation (Annovar) - novel transcripts discovery (Cufflinks) - predicting proteotypic peptides and creating ref/alt proteins fasta-database - integration of the results
    Downloads: 0 This Week
    Last Update:
    See Project
  • 9

    lr2rmats

    Long read to rMATS

    lr2rmats is a Snakemake-based light-weight pipeline which is designed to utilize both third-generation long-read and second-generation short-read RNA-seq data to generate an enhanced gene annotation file. The newly generated annotation file could be provided to rMATS for differential alternative splicing analysis. More information can be found at https://sourceforge.net/p/lr2rmats/wiki/Home/
    Downloads: 0 This Week
    Last Update:
    See Project
  • $300 Free Credits to Build on Google Cloud Icon
    $300 Free Credits to Build on Google Cloud

    New customers can spin up VMs, build with AI, and query data at no cost.

    Put your $300 in credit toward real workloads, then keep building with free monthly usage for 20+ products. No commitment and no charge until you upgrade.
    Start Free
  • 10
    SmartMuseum

    SmartMuseum

    Software for work with Corpus of Everyday life history Sources

    ...Corpuses of everyday life history sources are being collected in many museums and document archives. In this project, we consider the problem of creating software infrastructure for collaborative semantic annotation, information relation, and personalized access to corpus of everyday life history sources. Project financially supported from Department for Humanities of Russian Fund for Basic Research according to project # 16-01-12033. Authors: Vdovenko A., Marchenkov S., Petrina O., Korzun D.
    Downloads: 0 This Week
    Last Update:
    See Project
  • 11
    WikiSQL

    WikiSQL

    A large annotated semantic parsing corpus for developing NL interfaces

    A large crowd-sourced dataset for developing natural language interfaces for relational databases. WikiSQL is the dataset released along with our work Seq2SQL: Generating Structured Queries from Natural Language using Reinforcement Learning. Regarding tokenization and Stanza, when WikiSQL was written 3-years ago, it relied on Stanza, a CoreNLP python wrapper that has since been deprecated. If you'd still like to use the tokenizer, please use the docker image. We do not anticipate switching...
    Downloads: 0 This Week
    Last Update:
    See Project
  • 12

    MethyMer

    Design of specific primer combinations for bisulfite sequencing

    ...It also incorporates TCGA CpG methylation (microarrays) and gene expression (RNA-Seq) data, as well as methylation-expression correlation analysis results for 20 human cancer types. ENCODE genome regions annotation data are also integrated in MethyMer
    Downloads: 0 This Week
    Last Update:
    See Project
  • 13

    BioC

    We describe a simple XML format to share text documents and annotation

    A minimalist approach to share text documents and data annotations. Allows a large number of different annotations to be represented. Project files contain: - simple code to hold/read/write data and perform sample processing. - BioC-formatted corpora - BioC tools that work with BioC corpora BioC goals - simplicity - interoperability - broad use - reuse There should be little investment required to learn to use a format or a software module to process that format. We are...
    Leader badge
    Downloads: 11 This Week
    Last Update:
    See Project
  • 14
    ncPRO-seq

    ncPRO-seq

    Non-Coding RNA PROfiling from sRNA-seq

    ncPRO-seq is a tool for annotation and profiling of ncRNAs from smallRNA sequencing data. It aims to interrogate and perform detailed analysis on small RNAs derived from annotated non-coding regions in miRBase, piRBase, Rfam and repeatMasker, and regions defined by users. The ncPRO pipeline also has a module to identify regions significantly enriched with short reads that can not be classified as known ncRNA families. ############# Docker version : download and run Dockerfile (go in "Files" section) ############# GitHub : https://github.com/jbrayet/ncpro-seq
    Downloads: 0 This Week
    Last Update:
    See Project
  • 15
    Scaffold_Builder

    Scaffold_Builder

    Combining de novo and reference-guided assembly with Scaffold_builder

    ...Gaps are filled with N's and small overlaps are aligned with Needleman–Wunsch algorithm and the consensus created with IUPAC codes. Scaffold_builder can help in the assembly and annotation of genomes by revealing what is missing and allowing targeted sequencing to close those gaps. (c) Silva GG, Dutilh BE, Matthews TD, Elkins K, Schmieder R, Dinsdale EA, Edwards RA. Please cite: "Combining de novo and reference-guided assembly with Scaffold_builder", Source Code for Biology and Medicine 2013.
    Downloads: 0 This Week
    Last Update:
    See Project
  • 16
    ChIP-RNA-seqPRO

    ChIP-RNA-seqPRO

    ChIP-RNA-sequencing-processing (ChIP-RNA-seqPRO)

    ChIP-RNA-seqPRO: A strategy for identifying regions of epigenetic deregulation associated with aberrant transcript splicing and RNA-editing sites. Runnable python scripts packaged together with customized annotation libraries, demo data input and README guide. 9/26 : v1.1 Updated MAIN_IV to debug error thrown by python pandas no longer supporting 'subset'. This code will no longer be actively maintained/updated here. A cloud-based resource for comparative analysis of epigenetic, sequence variation, and expression datasets is now available. ...
    Downloads: 0 This Week
    Last Update:
    See Project
  • 17
    Square Genome Annotator
    Squere is a prokaryote genome annotation user-friendly software, with easy installer and graphical interface.
    Downloads: 0 This Week
    Last Update:
    See Project
  • 18

    mwetoolkit

    THIS PROJECT MIGRATED TO https://gitlab.com/mwetoolkit/mwetoolkit3/

    THIS PROJECT MIGRATED TO https://gitlab.com/mwetoolkit/mwetoolkit3/ The Multiword Expressions toolkit aids in the automatic identification and extraction of multiword units in running text. These include idioms (kick the bucket), noun compounds (cable car), phrasal verbs (take off, give up), etc. Even though it focuses on multiword expresisons, the framework is quite complete and can also be useful in any corpus-based study in computational linguistics. The mwetoolkit can be...
    Downloads: 0 This Week
    Last Update:
    See Project
  • 19
    mitoMaker

    mitoMaker

    mitoMaker - a mitochondria assembly and annotation script

    mitoMaker is a pipeline script developed to simplify the assembly and automatic annotation of mitochondrial genomes, based on raw NGS reads and an optional target reference. mitoMaker calls well known assemblers and algorithms, such as SOAPdenovo, MIRA and blast+ and parses their results providing easily readable outputs, such as FASTA, GENBANK, SEQUIN, PNG and others. General pipeline: 1-iterative De Novo assembly, with different k-mer values, trying to assemble a build that matches a target mitochondrial genome given. 2-searches for all mitochondrial gene features and circularization. 3-stores the best result found. 4-uses the best assembly as backbone for a reference based assembly, using MIRA and MITObim, trying to extend the mitogenome and close gaps. 5-annotates the best assembly, identifying the start and end position of each and every feature. 6-creates a folder with all the results (PNG, GENBANK, FASTA, SEQUIN, CAF, MAF and a stats logfile).
    Downloads: 0 This Week
    Last Update:
    See Project
  • 20
    ...IMPACT utilizes multi-reads in calling peaks and provides users with high-confidence peaks. In addition, IMPACT provides a completely integrated pipeline which produces downstream analysis results such as motif discovery and peak-to-gene annotation.
    Downloads: 0 This Week
    Last Update:
    See Project
  • 21
    SeqSelector

    SeqSelector

    Tools to select sequences for capture enrichment of next-gen libraries

    ...The scripts require no knowledge of programming, and can be applied to genome sequences of model or non-model species. We suggest a workflow in which genes of interest are first identified from previous studies and publicly available datasets of functional gene annotation. Once a list of candidate genes has been identified, their sequences are selected from the reference genome. These sequences are used as a query during a BLAST search of the unannotated genome of a non-model species, and then the corresponding sequences are returned, which can be used to design baits for hybridization-based sequence capture.
    Downloads: 0 This Week
    Last Update:
    See Project
  • 22
    Russian morphology tagger. Parses text(s) and output xml representation of text(s) with grammatical annotation.
    Downloads: 0 This Week
    Last Update:
    See Project
  • 23
    pyMantis
    pyMantis is a data-management system for (systems) biology build on the web2py framework. It features: tree based file explorer, relational db table wizzard with automated creation of user interfaces, internal and external access management, wiki, ..
    Downloads: 0 This Week
    Last Update:
    See Project
  • 24

    Genomic Binding Sites Analyser (BiSA)

    Genomic Region Archiving and Binding Sites Analysis (BiSA)

    ...BiSA can also annotate binding regions of interest with nearby genes. The results of overlap analysis can be imported into the Knowledge Base, allowing them to go into downstream analysis and independent annotation. A Venn diagram tool is also integrated into the software to allow users to visualize overlap results.
    Downloads: 0 This Week
    Last Update:
    See Project
  • 25
    Donatus is an on-going project consisting of Python, NLTK-based tools and grammars for deep parsing and syntactical annotation of Brazilian Portuguese corpora. It includes a user-friendly graphical user interface for building syntactic parsers with the NLTK, providing some additional functionalities.
    Downloads: 0 This Week
    Last Update:
    See Project