Hermes 4 70B (Nous Research): live API pricing, context window and the best open-source alternatives, tracked daily by olud.ai.
Prices update automatically — checked hourly against provider list prices.
Independent benchmark scores for Hermes 4 70B, measured by Artificial Analysis. Higher is better. Measurement mode: Non-reasoning.
Hermes 4 70B is an open-weight AI model by Nous Research. You can download and self-host it for free; the prices below are hosted-API list prices, tracked hourly, for when you prefer convenience over self-hosting.
Hermes 4 70B is an AI language model from Nous Research. It is open-weight: you can download it and run it on your own hardware, for free. It scores 6.9 on the Artificial Analysis intelligence index.
The weights are free and open — you can self-host Hermes 4 70B and pay nothing per token. If you prefer a hosted API, list prices are $0.13 per million input tokens and $0.4 per million output tokens.
Independent benchmarks from Artificial Analysis give it GPQA 49.1%, MMLU-Pro 66.4%, Humanity's Last Exam 3.6%, Long Context Reasoning 2%, LiveCodeBench 26.9%, SciCode 27.7%, AIME 2025 11.3%, IFBench 29%, τ²-Bench 21.6%, Terminal-Bench Hard 0%. It is particularly used for mathematical reasoning.
It generates about 89.9 tokens per second, with a median 0.6s delay before the first token. Measured independently by Artificial Analysis.
Yes. Hermes 4 70B has open weights, so you can download it and run it on your own GPU or server with tools like Ollama, vLLM or llama.cpp — with no per-token cost.