Exact checkpoint / Estimate-only preview
DeepSeek-R1-Distill-Qwen-32B (publisher BF16)
deepseek-ai/DeepSeek-R1-Distill-Qwen-32B · bfloat16 · safetensors
Revision 711ad2ea6aa40cfca18895e8aca02ab92df1a746
Metadata reviewed on 2026-09-27; next review 2026-10-27. Speed not measured. No tested rental configuration.
Model → workload → exact configuration
Find a GPU for DeepSeek-R1-Distill-Qwen-32B (publisher BF16)
Single-GPU text inference. Estimates remain unknown where exact hardware, runtime or charges lack evidence. NVGPU does not reserve stock or stop your instance.
Standalone memory and hourly-rate calculator
Estimate-only preview
Plan with DeepSeek-R1-Distill-Qwen-32B (publisher BF16)
One dense text model on one GPU. Input includes system instructions, conversation history and tool text. Output reserve is added to every simultaneous request.
Artifact and access evidence
DeepSeek derivative of Qwen/Qwen2.5-32B; distilled using DeepSeek-R1 generated examples according to the publisher card. Public ungated metadata and anonymous metadata access observed at review time. This does not establish individual access approval or verify a complete model download. Consult both linked license texts and preserve applicable notices.
Publisher terms: MIT · Upstream terms: Apache-2.0
Gated download: no. Authentication required: no. Access is subject to the linked terms.
Selected weights total 61.03 GiB on disk. Tokenizer, container, runtime workspace and retained data require additional space.
Pinned files and hashes
- model-00001-of-000008.safetensors · 8,792,578,462 bytes
SHA-256: 52addd84f7e83b6f5b54edfbcd11a9dc71ced13e63967e0e551fcdd071a09cac - model-00002-of-000008.safetensors · 8,776,906,899 bytes
SHA-256: 7401ea9ee01963efcae3bc43c34b17546236e30abfb356541e56df7ab2872d9a - model-00003-of-000008.safetensors · 8,776,906,927 bytes
SHA-256: 38b7e1968565c4334bae24082c262a337838d6520b590d4d0c3806df060d7a26 - model-00004-of-000008.safetensors · 8,776,906,927 bytes
SHA-256: 88f4a70419dc6c41779a88416baa0f4bad155e30c23dd8e90502bd98a32c9a49 - model-00005-of-000008.safetensors · 8,776,906,927 bytes
SHA-256: 45aee647ad5c930eb73896d507f0694b14218616cbe14b3f499eaab0eb5f1558 - model-00006-of-000008.safetensors · 8,776,906,927 bytes
SHA-256: bc3c1a143d5b7144a6cf58a1de95981562fe27354edb7bda87433b8da869740e - model-00007-of-000008.safetensors · 8,776,906,927 bytes
SHA-256: bb06d0ea82ab4a42820683158375e979a72bf3e88ebf0eb52080f4dabd9e54b6 - model-00008-of-000008.safetensors · 4,073,821,536 bytes
SHA-256: 1a6b6f41ddd7d1c07d4fdb9a626f66cb9a0fb3138c7d1116fc528255c6bb2b2b
Sources and completeness
- file-metadata · retrieved 2026-09-27
- configuration · retrieved 2026-09-27
- weight-index · retrieved 2026-09-27
- tokenizer · retrieved 2026-09-27
- publisher-card · retrieved 2026-09-27
- terms · retrieved 2026-09-27
- terms · retrieved 2026-09-27
- base_model_training_revision
- locally_verified_weight_checksums
- runtime_peak_measurements
- Publisher does not identify the exact base checkpoint training revision; parent repository attribution is known but its training commit remains unknown.
- SHA256 values are publisher-supplied LFS metadata; no full weights were downloaded or independently hashed.
- downloadBytes counts only selected safetensors shards; tensorBytes excludes serialization headers. Tokenizer, runtime image and temporary disk requirements belong to the runtime profile.
- maxPositionEmbeddings describes the publisher configuration limit, not a validated runtime context limit.
- Metadata review does not establish runtime support, device fit, speed, stock availability or complete cost.