DeepSeek Coder

DeepSeek-Coder is a series of code-specialized language models designed to generate, complete, and infill code (and mixed code + natural language) with high fluency in both English and Chinese. The models are trained from scratch on a massive corpus (~2 trillion tokens), of which about 87% is code and 13% is natural language. This dataset covers project-level code structure (not just line-by-line snippets), using a large context window (e.g. 16K) and a secondary fill-in-the-blank objective to encourage better contextual completions and infilling. Multiple sizes of the model are offered (e.g. 1B, 5.7B, 6.7B, 33B) so users can trade off inference cost vs capability. The repo provides model weights, documentation on training setup, evaluation results on common benchmarks (HumanEval, MultiPL-E, APPS, etc.), and inference tools.

Features

Multiple model sizes (1 B, 5.7 B, 6.7 B, 33 B) to suit different compute & use cases
Trained from scratch on ~2 trillion tokens, with 87% code and 13% natural language
Project-level context window (16K) and fill-in-the-blank objective for better infilling
Strong performance on code benchmarks (HumanEval, MultiPL-E, APPS, etc.)
Permissive license with “responsible downstream use” clause
Inference tooling and evaluation scripts for code generation and benchmarking

Project Samples

Project Activity

See All Activity >

License

MIT License

Follow DeepSeek Coder

DeepSeek Coder Web Site

Other Useful Business Software

Gen AI apps are built with MongoDB Atlas

Build gen AI apps with an all-in-one modern database: MongoDB Atlas

MongoDB Atlas provides built-in vector search and a flexible document model so developers can build, scale, and run gen AI apps without stitching together multiple databases. From LLM integration to semantic search, Atlas simplifies your AI architecture—and it’s free to get started.

Start Free

Rate This Project

User Reviews

Be the first to post a review of DeepSeek Coder!

Additional Project Details

Programming Language

Python

Related Categories

Python AI Models

Registered

2 days ago

Similar Business Software

DeepSeek-Coder-V2

DeepSeek-Coder-V2 is an open source code language model designed to excel in programming and mathematical reasoning tasks. It features a Mixture-of-Experts (MoE) architecture with 236 billion total parameters and 21 billion activated parameters per token, enabling efficient processing and high...

See Software
Baichuan-13B

Baichuan-13B is an open source and commercially available large-scale language model containing 13 billion parameters developed by Baichuan Intelligent following Baichuan -7B . It has achieved the best results of the same size on authoritative Chinese and English benchmarks. This release...

See Software
DeepSeek Coder

DeepSeek Coder is a cutting-edge software tool designed to revolutionize the landscape of data analysis and coding. By leveraging advanced machine learning algorithms and natural language processing capabilities, it empowers users to seamlessly integrate data querying, analysis, and...

See Software
Qwen-7B

Qwen-7B is the 7B-parameter version of the large language model series, Qwen (abbr. Tongyi Qianwen), proposed by Alibaba Cloud. Qwen-7B is a Transformer-based large language model, which is pretrained on a large volume of data, including web texts, books, codes, etc. Additionally, based on the...

See Software
DeepSeekMath

DeepSeekMath is a specialized 7B parameter language model developed by DeepSeek-AI, designed to push the boundaries of mathematical reasoning in open-source language models. It starts from the DeepSeek-Coder-v1.5 7B model and undergoes further pre-training with 120B math-related tokens sourced...

See Software
Qwen

Qwen LLM refers to a family of large language models (LLMs) developed by Alibaba Cloud's Damo Academy. These models are trained on a massive dataset of text and code, allowing them to understand and generate human-like text, translate languages, write different kinds of creative content, and...

See Software

Report inappropriate content

DeepSeek Coder

DeepSeek Coder: Let the Code Write Itself

Get an email when there's a new version of DeepSeek Coder

Features

Project Samples

Project Activity

Categories

License

Follow DeepSeek Coder

User Reviews

Additional Project Details

Programming Language

Related Categories

Registered