deepseek-r1-distill-llama-70b · DeepSeek
Provided by Groq, the DeepSeek-R1-Distill model is fine-tuned based on an open-source model, using samples generated by DeepSeek-R1. We have made slight modifications to their configurations and tokenizers. Please use our settings to run these models.
Provided by Groq, the DeepSeek-R1-Distill model is fine-tuned based on an open-source model, using samples generated by DeepSeek-R1. We have made slight modifications to their configurations and tokenizers. Please use our settings to run these models.
On AIHubMix, DeepSeek R1 Distill Llama 70B costs $0.8 per million input tokens and $1.6 per million output tokens.
DeepSeek R1 Distill Llama 70B accepts text input.
DeepSeek R1 Distill Llama 70B supports thinking. Per-protocol parameter support is listed in the capability table on this page.
DeepSeek R1 Distill Llama 70B is available through the AIHubMix unified API. The API is OpenAI-compatible: point your OpenAI SDK at https://aihubmix.com/v1, use your AIHubMix API key, and set the model name to deepseek-r1-distill-llama-70b — no other code changes needed.
DeepSeek R1 Distill Llama 70B is developed by DeepSeek. AIHubMix aggregates it alongside models from other providers behind one API and one bill.
DeepSeek R1 Distill Llama 70B was released on January 23, 2025 by DeepSeek.
DeepSeek-V4.1-Flash model official release. This is the smallest model in DeepSeek’s new…
DeepSeek-V4-Flash-0731(deepseek-v4-flash-0731) is an open-source MoE large language model…
DeepSeek’s officially released new multimodal visual-understanding model…
DeepSeek V4 Pro 0813 is DeepSeek’s high-performance general-purpose reasoning and agent…
DeepSeek V4 Flash 0731 Fast is a high-speed deployment of DeepSeek’s agentic model…
(This model currently points to the older 0423 version; if you need to request the latest…
Use DeepSeek R1 Distill Llama 70B via the AIHubMix unified API — one interface for every major LLM.