Nous Hermes 3 70B

Text & chatUnavailable
by Together AIModel ID: hermes-3-70b

Llama-3.1-70B fine-tune from Nous Research with strong tool/agent capabilities and uncensored alignment.

Status
Unavailable
Context
131.072 tokens
Max. output
4.096 tokens
Input โ†’ output
Text โ†’ Text
Developer
Together AI
Updated
June 25, 2026

Nous Hermes 3 70B is currently unavailable

Currently unavailable: this model has been deactivated.

You can still read the details on this page. Pick one of the available alternatives below to run a comparable model right away.

Go to alternatives
01

Comparable models

All in this category
  • Anthropic's model for the most demanding reasoning and long-horizon agentic work. 1M-token context window, up to 128K output tokens, adaptive thinking that is always on.

    US$12.00/1M in

  • The most capable model of Anthropic's Opus 4 series. State of the art on long-horizon agentic work, coding and knowledge tasks, with a 1M-token context window at standard pricing.

    US$6.00/1M in

  • Anthropic's current Opus model for long-running agentic coding and knowledge work. 1M-token context window, up to 128K output tokens, adaptive thinking that is always on.

    US$4.80/1M in

02

Playground

Try Nous Hermes 3 70B

Chat

Currently unavailable

Currently unavailable: this model has been deactivated.

The playground is disabled. You can find comparable models in the same category: Browse alternatives

Try Nous Hermes 3 70B

Send a message. The answer arrives in full once the model is done (no streaming).

System prompt
Max. answer length (tokens)

This run

No price โ€“ currently unavailable.

New here?

5 free credits (US$0.05) when you sign up with Google

Usable 24 hours after sign-up, up to 5 runs per day and at most 2 credits per run. Other sign-in methods start without credits.

03

About Nous Hermes 3 70B

TL;DRAs of June 25, 2026

Nous Hermes 3 70B is a model by Together AI in the Text & chat category. Nous Hermes 3 70B is currently not available on Railwail. The context window holds 131.072 tokens, and one response can be up to 4.096 tokens long.

Background

About Nous Research

Founded 2023 ยท San Francisco, USA (distributed)

Nous Research is a community-driven open-source AI collective founded in 2023, co-led by Karan 'Teknium' Malhotra, Jeffrey Quesnelle and Bowen Peng with a distributed contributor base. Nous focuses on uncensored, steerable, character-rich fine-tunes of open base models. The Hermes line (Hermes, OpenHermes, Hermes 2, Hermes 2.5, Hermes 3) is its flagship instruction-following family. Hermes 3 70B is the mid-sized variant of the August 2024 Hermes 3 release, fine-tuned from Llama 3.1 70B for steerable behaviour with native tool-use and scratchpad tags โ€” the sweet spot for organisations that can run a 70B model on a small GPU cluster.

Visit Nous Research

Architecture

Decoder-only Transformer (Llama 3.1 architecture)

Hermes 3 70B is a full-parameter supervised fine-tune of Meta's Llama 3.1 70B base. The architecture is identical to Llama 3.1 70B: 80 layers, 8,192 hidden size, 64-head grouped-query attention with 8 KV heads, RoPE positional embeddings with Llama 3 scaling (128K context), SwiGLU activations, and the 128,000-token Llama 3 BPE tokeniser. The fine-tune mixes role-play, function calling, code, math, RAG, agent traces and creative writing โ€” largely Nous-curated synthetic data distilled from larger models. The model uses ChatML formatting with native `<tool_call>` JSON-schema tags and `<scratchpad>` chain-of-thought tags. A light DPO preference-optimisation stage was applied (unlike the 405B which is SFT-only). Released August 2024 under the Llama 3.1 Community License.

Parameters
70B (dense)
Context
128.000 tokens

Capabilities

  • Full-parameter fine-tune of Llama 3.1 70B with light DPO
  • Excellent system-prompt steerability for agents and characters
  • Native ChatML `<tool_call>` and `<scratchpad>` tags
  • More tractable to self-host than the 405B (~140GB FP16, ~35GB INT4)
  • Strong instruction following and creative writing
  • 128K context inherited from Llama 3.1
  • Comparable to Llama 3.1 70B Instruct on benchmarks with friendlier tool-use format
  • Best for: self-hosted assistants, tool-using agents, creative platforms, fine-tuning base for community projects.

Training & license

Supervised fine-tuning on ~390M instruction tokens across ~2.5M examples covering role-play, function calling, code, math, RAG, agent traces and creative writing โ€” largely Nous-curated synthetic data. A light DPO preference-optimisation pass was applied. Base knowledge cutoff December 2023.

License: Llama 3.1 Community License. Commercial use permitted, but services with >700M monthly active users require a separate Meta license. Meta's Acceptable Use Policy applies to all derivatives.

Safety testing: Nous explicitly trades RLHF-style safety guardrails for steerability and creative latitude. Operators deploying for consumer products should add their own safety layer. No formal third-party red-team report.

Known limitations

  • Reduced safety guardrails versus Meta's Llama 3.1 70B Instruct
  • Less capable than the 405B variant on hard reasoning, math and code
  • No vision modality
  • Knowledge cutoff inherited from Llama 3.1 (December 2023)
  • Llama 3.1 license restricts services with >700M MAU
04

Pricing

Currently unavailable: this model has been deactivated. There is no price for this model at the moment, so it cannot be run.

05

API

Call Nous Hermes 3 70B with your Railwail API key. Use this model ID in the request:

Currently unavailable

The model has no verified price or is deactivated; API calls are refused.

06

Specifications

Model ID
hermes-3-70b
Developer
Together AI
Category
Text & chat
Input
Text
Output
Text
Context window
131.072 tokens
Max. output
4.096 tokens
Model size
70B (dense)
License
Llama 3.1 Community License. Commercial use permitted, but services with >700M monthly active users require a separate Meta license. Meta's Acceptable Use Policy applies to all derivatives.
Catalog entry updated
June 25, 2026

Tags

  • nous
  • open-weights
  • tools
  • roleplay
  • pricing-tbd
07

Use cases

What it is used for

  • Self-hosted assistants on 2x H100 or single A100 80GB
  • Tool-using agents with ChatML format
  • Creative writing and role-play platforms
  • Fine-tuning base for downstream community models
  • Customisable persona-driven chat
  • Function-calling integrations with LangChain / DSPy
08

Frequently asked questions

What is Nous Hermes 3 70B?

Nous Hermes 3 70B is a model by Together AI in the Text & chat category. It is listed on Railwail but cannot be run at the moment.

How much does Nous Hermes 3 70B cost on Railwail?

Nous Hermes 3 70B cannot be run on Railwail at the moment, so there is no current price. Available alternatives with prices are listed further down this page.

What is the context window of Nous Hermes 3 70B?

The context window of Nous Hermes 3 70B holds 131.072 tokens. One response can be up to 4.096 tokens long.

How fast is Nous Hermes 3 70B?

There are not enough measured runs of Nous Hermes 3 70B on Railwail yet to state a run time. It depends on the input, the settings and the load at the provider.

Is Nous Hermes 3 70B better than Claude Fable 5.1?

That depends on the task. Nous Hermes 3 70B (Together AI) and Claude Fable 5.1 (Anthropic) are both models in the Text & chat category. The comparison page shows their prices and specifications side by side.

Compare Nous Hermes 3 70B and Claude Fable 5.1

Can I use Nous Hermes 3 70B right now?

Currently unavailable: this model has been deactivated. The page stays online; available alternatives from the same category are listed further down.

All models through one API

One API key for every model on Railwail. Usage is charged from prepaid credits, 1 credit = US$0.01.