Revision: 916b56a44061fd5cd7d6a8fb632557ed4f724f60
MIT; review upstream model terms. Public checkpoint. Use the pinned revision and tokenizer from the shared model registry.
Reviewed guide · no hardware test recorded
Plan a single-GPU DeepSeek-R1-Distill-Qwen-7B BF16 session with the existing vLLM preview profile.
chat-deepseek-7b · reviewed 2026-09-27 · review due 2026-10-27
Model shard sizes are recorded in the model registry; total environment download and host requirements are unknown. Use the shared model preview for conditional memory arithmetic.
Changing task dimensions invalidates a tested-fit claim for a different task. This version has 0 recorded successful runs.
Revision: 916b56a44061fd5cd7d6a8fb632557ed4f724f60
MIT; review upstream model terms. Public checkpoint. Use the pinned revision and tokenizer from the shared model registry.
Shared runtime: vllm-0.11.0-qwen2-bf16-single-gpu-preview-v1. Pinned image: vllm/vllm-openai@sha256:d8d39b59e909d2378ac4feeb191f7e7b6f1342477dc66b7c47cec89e9985ad8a.
Shared model variant: deepseek-r1-distill-qwen-7b-bf16@916b56a44061fd5cd7d6a8fb632557ed4f724f60. Open memory arithmetic preview
On-demand GPU Pod; verify driver and image support before renting.
Enter your own cold-download, loading and warmup allowance. Warm restart times have not been measured.
NVGPU does not provision, stop or delete resources.
Official source: Pinned model identity; metadata is shared with the model directory.
Official source: Manual Pod setup and lifecycle actions.