DeepSeek Coder 33B Instruct (GGUF)
deepseek-coder-33b-instructQuantized GGUF build of DeepSeek's 33B code model, trained on roughly 2T tokens that are about 87 percent code. Designed for repository-level completion and project-aware generation thanks to a 16k context window. Runs on Replicate as a per-call endpoint.
- Price
- โ US$0.0030/run
- Context
- 16,384 tokens
- Max. output
- 4,096 tokens
- Input โ output
- Text โ Text
- Developer
- Community
- Updated
- September 23, 2026
Playground
Try DeepSeek Coder 33B Instruct (GGUF)
Input & output
This run
about US$0.003 ยท 0.3 credits
US$0.009 (0.9 credits) are reserved at the start; the actual GPU time is billed.
New here?
10 free credits (US$0.10) when you sign up with Google
Usable 24 hours after sign-up, up to 5 runs per day and at most 2 credits per run. Other sign-in methods start without credits. Enough for 11 runs of this model.
Examples
Prompt
please create a rust enum called prediction status, with three variants starting, in progress and completed. Please only include valid rust code, do not include any commentary or explanations.
Response
```rust enum PredictionStatus { Starting, InProgress, Completed, } ```
Settings
system_prompt: You are an AI programming assistant, utilizing the Deepseek Code model, developed by Deepseek Company, and you only answer questions related to computer science. For politically sensitive questions, security and privacy issues, and other non-computer science questions, you will refuse to answer.
About DeepSeek Coder 33B Instruct (GGUF)
DeepSeek Coder 33B Instruct (GGUF) is a model by Community in the Code category. On Railwail, DeepSeek Coder 33B Instruct (GGUF) costs โ US$0.0030 per run. The context window holds 16,384 tokens, and one response can be up to 4,096 tokens long.
Pricing
| Typical run (โ 3 s on L40S) | US$0.0030 per run |
|---|---|
| GPU time (L40S) | US$0.00117 per GPU second |
- Billed by the GPU time the run actually takes. When the run starts, 3ร the typical price is reserved from your balance and settled afterwards.
- 1 credit = US$0.01
Cost calculator
Price calculator
Typical according to the provider: about 2.6 s
Total
US$0.30
30 credits
Per run
US$0.003 ยท 0.3 credits
Billed by the actual GPU time; this is an estimate.
API
No verified API example
The public API passes a different input format than this model needs. Use the playground above.
Specifications
- Model ID
deepseek-coder-33b-instruct- Developer
- Community
- Category
- Code
- Input
- Text
- Output
- Text
- Context window
- 16,384 tokens
- Max. output
- 4,096 tokens
- Billing
- By usage (tokens or GPU time)
- Catalog entry updated
- September 23, 2026
Input parameters
Inputs and settings from the model's input schema. The example in the API section shows which of them the API accepts.
promptrequiredCoding prompt
Type: TextDefault: โAllowed values: โtemperatureType: NumberDefault:0.2Allowed values: 0 to 2max_new_tokensType: IntegerDefault:1024Allowed values: 1 to 4,096
Tags
- deepseek
- coding
- instruct
- gguf
- open-weights
- replicate
Use cases
Frequently asked questions
What is DeepSeek Coder 33B Instruct (GGUF)?
DeepSeek Coder 33B Instruct (GGUF) is a model by Community in the Code category.
How much does DeepSeek Coder 33B Instruct (GGUF) cost on Railwail?
On Railwail, DeepSeek Coder 33B Instruct (GGUF) costs โ US$0.0030 per run. You are charged for what each request actually uses. Usage is paid from prepaid credits; 1 credit equals US$0.01.
What is the context window of DeepSeek Coder 33B Instruct (GGUF)?
The context window of DeepSeek Coder 33B Instruct (GGUF) holds 16,384 tokens. One response can be up to 4,096 tokens long.
How fast is DeepSeek Coder 33B Instruct (GGUF)?
There are not enough measured runs of DeepSeek Coder 33B Instruct (GGUF) on Railwail yet to state a run time. It depends on the input, the settings and the load at the provider.
Is DeepSeek Coder 33B Instruct (GGUF) better than Code Llama 13B Instruct?
That depends on the task. DeepSeek Coder 33B Instruct (GGUF) (Community) and Code Llama 13B Instruct (Meta) are both models in the Code category. The comparison page shows their prices and specifications side by side.
Compare DeepSeek Coder 33B Instruct (GGUF) and Code Llama 13B InstructComparable models
All in this categoryMeta's 13B Code Llama tuned for instruction following. A faster mid-size option for code generation and completion, supporting infilling for inserting code at a cursor position. Served on Replicate per call.
โ US$0.0065/run
117 % more expensive per unit
Compare DeepSeek Coder 33B Instruct (GGUF) vs. Code Llama 13B InstructMeta's 34B Code Llama tuned for instruction following. A balance of size and quality for code generation, completion, and explanation, with strong coverage of Python, JavaScript, and other common languages. Runs on Replicate per call.
โ US$0.0408/run
1,260 % more expensive per unit
Compare DeepSeek Coder 33B Instruct (GGUF) vs. Code Llama 34B InstructMeta's largest Code Llama, a 70B Llama-2 derivative specialized for programming and tuned to follow instructions in chat form. Handles code generation, completion, and explanation across common languages. Served on Replicate as a per-call endpoint.
โ US$0.0504/run
1,580 % more expensive per unit
Compare DeepSeek Coder 33B Instruct (GGUF) vs. Code Llama 70B Instruct
All models through one API
One API key for every model on Railwail. Usage is charged from prepaid credits, 1 credit = US$0.01.