DeepSeek Coder 33B Instruct (GGUF)

CodeAvailable
by CommunityModel ID: deepseek-coder-33b-instruct

Quantized GGUF build of DeepSeek's 33B code model, trained on roughly 2T tokens that are about 87 percent code. Designed for repository-level completion and project-aware generation thanks to a 16k context window. Runs on Replicate as a per-call endpoint.

Price
โ‰ˆ $0.0030/run
Context
16,384 tokens
Max. output
4,096 tokens
Input โ†’ output
Text โ†’ Text
Developer
Community
Updated
September 23, 2026
01

Playground

Try DeepSeek Coder 33B Instruct (GGUF)

Input & output

โ‰ˆ $0.0030/run
Try DeepSeek Coder 33B Instruct (GGUF)

Coding prompt

Advanced settings (2)
Output
The answer appears here.

This run

about $0.003 ยท 0.3 credits

$0.009 (0.9 credits) are reserved at the start; the actual GPU time is billed.

New here?

10 free credits ($0.10) when you sign up with Google

Usable 24 hours after sign-up, up to 5 runs per day and at most 2 credits per run. Other sign-in methods start without credits. Enough for 11 runs of this model.

02

Examples

Real outputs from the public examples of this model on Replicate, with the prompt and settings that produced them. They were not generated live on this page.
  • Prompt

    please create a rust enum called prediction status, with three variants starting, in progress and completed. Please only include valid rust code, do not include any commentary or explanations.

    Response

    ```rust enum PredictionStatus { Starting, InProgress, Completed, } ```

    Settings

    system_prompt: You are an AI programming assistant, utilizing the Deepseek Code model, developed by Deepseek Company, and you only answer questions related to computer science. For politically sensitive questions, security and privacy issues, and other non-computer science questions, you will refuse to answer.

03

About DeepSeek Coder 33B Instruct (GGUF)

TL;DRAs of September 23, 2026

DeepSeek Coder 33B Instruct (GGUF) is a model by Community in the Code category. On Railwail, DeepSeek Coder 33B Instruct (GGUF) costs โ‰ˆ $0.0030 per run. The context window holds 16,384 tokens, and one response can be up to 4,096 tokens long.

DeepSeek-Coder-33B-Instruct was one of the strongest open code models at release, built by DeepSeek and trained from scratch on a code-dominated corpus. This Replicate deployment uses a GGUF-quantized weight file for efficient inference. Good for code completion, generation, and refactoring across many languages.
04

Pricing

Prices in US dollars. Usage is charged from prepaid credits.
Typical run (โ‰ˆ 3 s on L40S)$0.0030 per run
GPU time (L40S)$0.00117 per GPU second
  • Billed by the GPU time the run actually takes. When the run starts, 3ร— the typical price is reserved from your balance and settled afterwards.
  • 1 credit = $0.01

Cost calculator

Price calculator

s

Typical according to the provider: about 2.6 s

Total

$0.30

30 credits

Per run

$0.003 ยท 0.3 credits

Billed by the actual GPU time; this is an estimate.

05

API

Call DeepSeek Coder 33B Instruct (GGUF) with your Railwail API key. Use this model ID in the request:
deepseek-coder-33b-instructAPI documentationGet an API key

No verified API example

The public API passes a different input format than this model needs. Use the playground above.

06

Specifications

Model ID
deepseek-coder-33b-instruct
Developer
Community
Category
Code
Input
Text
Output
Text
Context window
16,384 tokens
Max. output
4,096 tokens
Billing
By usage (tokens or GPU time)
Catalog entry updated
September 23, 2026

Input parameters

Inputs and settings from the model's input schema. The example in the API section shows which of them the API accepts.

  • promptrequired

    Coding prompt

    Type: Text
    Default: โ€“
    Allowed values: โ€“
  • temperature
    Type: Number
    Default: 0.2
    Allowed values: 0 to 2
  • max_new_tokens
    Type: Integer
    Default: 1024
    Allowed values: 1 to 4,096

Tags

  • deepseek
  • coding
  • instruct
  • gguf
  • open-weights
  • replicate
07

Use cases

08

Frequently asked questions

What is DeepSeek Coder 33B Instruct (GGUF)?

DeepSeek Coder 33B Instruct (GGUF) is a model by Community in the Code category.

How much does DeepSeek Coder 33B Instruct (GGUF) cost on Railwail?

On Railwail, DeepSeek Coder 33B Instruct (GGUF) costs โ‰ˆ $0.0030 per run. You are charged for what each request actually uses. Usage is paid from prepaid credits; 1 credit equals $0.01.

What is the context window of DeepSeek Coder 33B Instruct (GGUF)?

The context window of DeepSeek Coder 33B Instruct (GGUF) holds 16,384 tokens. One response can be up to 4,096 tokens long.

How fast is DeepSeek Coder 33B Instruct (GGUF)?

There are not enough measured runs of DeepSeek Coder 33B Instruct (GGUF) on Railwail yet to state a run time. It depends on the input, the settings and the load at the provider.

Is DeepSeek Coder 33B Instruct (GGUF) better than Code Llama 13B Instruct?

That depends on the task. DeepSeek Coder 33B Instruct (GGUF) (Community) and Code Llama 13B Instruct (Meta) are both models in the Code category. The comparison page shows their prices and specifications side by side.

Compare DeepSeek Coder 33B Instruct (GGUF) and Code Llama 13B Instruct
09

Comparable models

All in this category

All models through one API

One API key for every model on Railwail. Usage is charged from prepaid credits, 1 credit = $0.01.