Code Models

AI-powered coding assistants for development

Modele do generowania kodu — autouzupełnianie, review i refaktoryzacja

Modele do generowania kodu to duże modele językowe trenowane lub fine-tunowane konkretnie na kodzie źródłowym. Zasilają autouzupełnianie w IDE, PR review, zautomatyzowaną refaktoryzację, generowanie testów i tłumaczenie cross-language. Po model kodu — zamiast ogólnego modelu tekstowego — sięgasz, gdy chcesz silniejszej poprawności w zadaniach programistycznych oraz ustrukturyzowanego outputu (diffy, JSON), który dobrze współpracuje z tooling developerskim.

21 models available

Codestral

CodeMistral AI
NewPopular

Mistral's code-specialized model. Optimized for code generation, completion, and understanding across 80+ languages.

Free1.5s
codingfastmultilanguage

Code Llama 13B Instruct

CodeMeta

Meta's 13B Code Llama tuned for instruction following. A faster mid-size option for code generation and completion, supporting infilling for inserting code at a cursor position. Served on Replicate per call.

€1.00
metacode-llamacoding

Code Llama 34B Instruct

CodeMeta

Meta's 34B Code Llama tuned for instruction following. A balance of size and quality for code generation, completion, and explanation, with strong coverage of Python, JavaScript, and other common languages. Runs on Replicate per call.

€2.00
metacode-llamacoding

Code Llama 70B Instruct

CodeMeta

Meta's largest Code Llama, a 70B Llama-2 derivative specialized for programming and tuned to follow instructions in chat form. Handles code generation, completion, and explanation across common languages. Served on Replicate as a per-call endpoint.

€3.00
metacode-llamacoding

Code Llama 7B Instruct

CodeMeta

Meta's smallest Code Llama at 7B parameters, tuned for instruction following. The cheapest and fastest member of the family for quick code generation, completion, and infilling. Served on Replicate per call.

€1.00
metacode-llamacoding

CodeGen 350M Mono

Codehuggingface

350M autoregressive code generation model from Salesforce, the smallest of the original CodeGen family. The mono variant was further trained on Python so it is well suited for short Python completions and program synthesis from a natural-language or code prompt.

Free
codehuggingfacesalesforce

DeepSeek Coder 1.3B Instruct

Codehuggingface

1.3B instruction-tuned code model from DeepSeek, trained on 2 trillion tokens of code and natural language across 87 languages with a 16k context window. One of the strongest tiny coders for its size, handling generation, completion and short coding instructions.

Free
codehuggingfacedeepseek

DeepSeek Coder 33B Instruct (GGUF)

CodeReplicate

Quantized GGUF build of DeepSeek's 33B code model, trained on roughly 2T tokens that are about 87 percent code. Designed for repository-level completion and project-aware generation thanks to a 16k context window. Runs on Replicate as a per-call endpoint.

€2.00
deepseekcodinginstruct

DeepSeek Coder V2

CodeDeepSeek

DeepSeek's specialized coding model. Excellent at code generation, debugging, and explanation.

Free2.0s
codingaffordable

Granite Code 20B

CodeReplicate

IBM Granite 20B Code Instruct. Larger Granite code model balancing quality and inference cost for enterprise CI/CD code-review automation.

€0.006
replicatecode-generationibm

Granite Code 8B

CodeReplicate

IBM Granite 8B Code Instruct. Trained on permissively-licensed code, strong on multi-language code completion and instruction-following.

€0.004
replicatecode-generationibm

Grok Build 0.1

CodexAI

xAI's Grok coding-focused model. Tuned for code generation and software development tasks with a 256k token context window for working over large codebases.

Free
xaigrokcode

Magicoder S CL 7B

CodeCommunity

UIUC Magicoder S CL 7B. CodeLlama-7B fine-tuned with OSS-Instruct synthetic data. Strong HumanEval Plus and MBPP Plus performance per parameter.

€0.003
replicatecode-generationopen-weights

Phind CodeLlama 34B v2

CodeReplicate

Phind CodeLlama 34B v2. Highly tuned CodeLlama variant focused on retrieval-augmented developer assistant workflows.

€0.009
replicatecode-generationphind

Qwen2.5-Coder 32B Instruct

Codehuggingface

Alibaba's largest open Qwen2.5-Coder model. Trained on a code-heavy corpus, it matches or beats much larger general models on code generation and repair benchmarks like HumanEval and MBPP, and supports over 40 programming languages with fill-in-the-middle completion.

Free
qwenalibabacoding

Qwen2.5-Coder 7B Instruct

Codehuggingface

The 7B instruct member of Alibaba's Qwen2.5-Coder family. A lighter, faster option for code completion, generation, and bug fixing across 40+ languages, with a 128k context and fill-in-the-middle support. Good price-to-quality balance for everyday coding tasks.

Free
qwenalibabacoding

Replit Code v1 3B

CodeReplicate

Replit's 3B code-completion model, trained on a permissively licensed code subset of the Stack across 20 programming languages. Built for low-latency autocomplete rather than chat. Served on Replicate per call.

€1.00
replitcodingcompletion

Replit Code v1.5 3B

Codehuggingface

3B code completion model from Replit trained on roughly 1 trillion tokens of permissively licensed code across 30 programming languages, with a 4k context window. Designed for autocomplete-style code generation and fill-in-the-middle.

Free
codehuggingfacereplit

Stable Code Instruct 3B

Codehuggingface

Instruction-tuned 3B code model from Stability AI, fine-tuned from stable-code-3b for chat-style coding tasks. Handles code generation, explanation and fix-up across multiple languages and was competitive with larger code models on benchmarks at release.

Free
codehuggingfacestability-ai

StarCoder2 15B

CodeCommunity

BigCode StarCoder2 15B code-generation flagship. Trained on 4T tokens of Stack v2 data with grouped-query attention and 16k context.

€0.005
replicatecode-generationbigcode

WizardCoder 33B

CodeCommunity

WizardLM WizardCoder 33B v1.1. Evol-Instruct fine-tune of DeepSeek-Coder-33B with strong code-generation benchmark performance.

€0.009
replicatecode-generationwizardlm

Top code models picks

Hand-picked across four common criteria — resolved against the live catalog so the picks track price and performance changes.

Najlepszy ogólnie
Codestral

Mistral's code-specialized model. Optimized for code generation, completion, and understanding across 80+ languages.

Learn more
Najtańszy
CodeGen 350M Mono

350M autoregressive code generation model from Salesforce, the smallest of the original CodeGen family. The mono variant was further trained on Python so it is well suited for short Python completions and program synthesis from a natural-language or code prompt.

Learn more
Najdłuższy kontekst
Codestral

Mistral's code-specialized model. Optimized for code generation, completion, and understanding across 80+ languages.

Learn more
Najszybszy
Codestral

Mistral's code-specialized model. Optimized for code generation, completion, and understanding across 80+ languages.

Learn more

Cennik w generowaniu kodu idzie tym samym modelem per-token co ogólny tekst. Flagshipowe modele kodu (GPT-5 Codex, Claude 4.6 Sonnet, Codestral) kosztują €1-€10 za milion tokenów wejściowych; tiery budżetowe (Codestral Mamba, DeepSeek Coder, Qwen Coder) kosztują €0,05-€0,50 za milion. Pojedyncze żądanie autouzupełniania w IDE rzadko mieści więcej niż kilka tysięcy tokenów wejściowych, więc koszt na wywołanie to ułamek centa. Rachunki rosną, gdy wypuszczasz agentów, którzy sami się reprompują dziesiątki razy w jednym zadaniu.

Trójkąt kompromisu to poprawność, prędkość i kontekst. Flagshipy rozwiązują trudniejsze problemy i pewniej trzymają się konwencji projektu, ale odpowiadają z prędkością 30-80 tokenów/sekundę, co czuje się wolno wewnątrz ciasnej pętli autouzupełniania. Szybkie budżetowe modele (Codestral Mamba, GPT-5 Mini) strumieniują z 200+ tokenami/sekundę i czują się natywnie w edytorze. Dla zadań batch (refaktoryzacja całego repo, generowanie testów do pięćdziesięciu plików) wygrywa poprawność flagshipa. Dla ciasnych pętli autouzupełniania wygrywa szybki tier.

Uwaga na kontekst cross-file: większość pętli autouzupełniania wysyła tylko bieżący plik. Dla prawdziwie codebase-aware refaktoryzacji potrzebna jest warstwa retrievalu, która wciąga powiązane pliki do promptu. Narzędzia jak Cursor i Continue robią to automatycznie; jeśli budujesz własne, najpierw zembeduj codebase i pobieraj 5-10 najbardziej trafnych plików na żądanie.

Uwaga na skażenie licencyjne: kilka modeli kodu open-weights było trenowanych tylko na kodzie z permisywnymi licencjami; inne zgarnęły kod GPL z niejasnymi warunkami redystrybucji. Jeśli wypuszczasz wygenerowany kod w produkcie closed-source, preferuj modele komercyjne z jawnymi gwarancjami dotyczącymi licencji kodu.

Top picks powyżej obejmują najbardziej poprawnego flagshipa, najtańszego konia roboczego, model o najdłuższym kontekście i najszybszą opcję autouzupełniania.

Frequently asked questions

Start Building with AI

Access all models through a single API. Get free credits when you sign up — no credit card required.