{"umans-qwen3.8-27b-lab":{"display_name":"Umans Qwen3.8 27B (lab)","provider":"Qwen","retired_on":"2026-08-21","stage":"playground","note":"playground experiment; the Qwen3.8 27B lab window ran Aug 18 to 21, 2026","description":"Qwen3.8 27B as a Labs experiment, open for a short test window: temporary, not a permanent id. Served from Qwen's official FP8 checkpoint: a compact dense model with native image and video understanding, built for coding and long-horizon agentic tasks on a 256K context window. It thinks by default at xhigh effort; reasoning can be tuned down (low, medium) or turned off (none). Access is seat-gated through the Labs page while an experiment is live. It is offered at limited capacity and low availability, so expect it to be flaky and to go down under load: crash it, give it a moment, and try again. For production work we recommend umans-coder or umans-deepseek-v4-pro-0813.","weights":{"precision":"fp8","hf_url":"https://huggingface.co/Qwen/Qwen3.8-27B-FP8"}},"umans-deepseek-v4-pro-0813-lab":{"display_name":"Umans DeepSeek V4 Pro 0813 (lab)","provider":"DeepSeek","retired_on":"2026-08-15","replacement":"umans-deepseek-v4-pro-0813","stage":"playground","note":"playground experiment; the V4 Pro 0813 pre-release window ran Aug 13 to 15, 2026; its metrics carry on the released model's status page as its pre-release period","description":"DeepSeek V4 Pro as a Labs experiment, open for a short test window: temporary, not a permanent id. Served from DeepSeek's pre-release 0813 checkpoint: the flagship coding and reasoning MoE, ahead of its official release. Reasoning has three modes: non-think (none), think high (high, the default) and think max (max). Access is seat-gated through the Labs page while an experiment is live. It is offered at limited capacity and low availability, so expect it to be flaky and to go down under load: crash it, give it a moment, and try again. For production work we recommend umans-coder or umans-glm-5.2.","weights":{"precision":"full","hf_url":"https://huggingface.co/deepseek-ai/DeepSeek-V4-Pro-0813"}},"umans-deepseek-v4-flash-0731-vision-lab":{"display_name":"Umans DeepSeek V4 Flash Vision (lab)","provider":"DeepSeek","retired_on":"2026-08-13","stage":"playground","note":"playground experiment; the vision V4 Flash lab window ran Aug 11 to 13, 2026","description":"DeepSeek V4 Flash Vision as a Labs experiment, open for a short test window: temporary, not a permanent id. Our own V4 Flash with vision added: the same performance and speed as DeepSeek V4 Flash (284B total, 13B active, 1M-token context), plus the ability to read images. Reasoning has four modes: non-think (none), think low (low, the default), think high (high) and think max (max). Access is seat-gated through the Labs page while an experiment is live. It is offered at limited capacity and low availability, and it is served on our own GPU infrastructure, so expect it to be flaky and to go down under load: crash it, give it a moment, and try again.","weights":{"precision":"full","hf_url":"https://huggingface.co/umans-ai/DeepSeek-V4-Flash-0731-Vision"}},"umans-deepseek-v4-flash-0731-lab":{"display_name":"Umans DeepSeek V4 Flash (lab)","provider":"DeepSeek","retired_on":"2026-08-11","stage":"playground","note":"playground experiment; the text V4 Flash lab window ended when the vision lab opened","description":"DeepSeek V4 Flash as a Labs experiment, open for a short test window: temporary, not a permanent id. DeepSeek's fast agentic coding MoE (284B total, 13B active), served from the official 0731 release, on a 1M-token context. Reasoning has four modes: non-think (none), think low (low, the default), think high (high) and think max (max). Access is seat-gated through the Labs page while an experiment is live. It is offered at limited capacity and availability, so expect it to be flaky and to go down under load: crash it, give it a moment, and try again. When the window ends, the model keeps serving as the pay-per-token umans-deepseek-v4-flash-0731.","weights":{"precision":"full","hf_url":"https://huggingface.co/deepseek-ai/DeepSeek-V4-Flash-0731"}},"umans-deepseek-v4-pro-dspark":{"display_name":"Umans DeepSeek V4 Pro DSpark","provider":"DeepSeek","retired_on":"2026-07-16","stage":"playground","note":"playground experiment; the test window ran Jul 14 to 16, 2026","report_url":"https://blog.umans.ai/blog/deepseek-v4-pro-dspark-the-architecture-is-ready/","description":"DeepSeek V4 Pro as a Labs experiment, open for a short test window (July 14 to 18, 2026): temporary, not a permanent model. DeepSeek's flagship coding and reasoning MoE, served from the original weights with speculative decoding for speed. Reasoning has three modes: non-think (none), think high (high, the default) and think max (max). Access is seat-gated through the Labs page while an experiment is live. It is offered at limited capacity and availability, so expect it to be flaky and to go down under load: crash it, give it a moment, and try again. For production work we recommend umans-coder or umans-glm-5.2.","weights":{"precision":"full","hf_url":"https://huggingface.co/deepseek-ai/DeepSeek-V4-Pro-DSpark"}},"umans-glm-5.2-nvfp4":{"display_name":"Umans GLM 5.2 NVFP4","provider":"GLM","retired_on":"2026-07-02","replacement":"umans-glm-5.2","stage":"playground","note":"playground experiment; the test window ran Jun 29 to Jul 2, 2026","report_url":"https://blog.umans.ai/blog/glm-5-2-nvfp4-not-worth-serving/","description":"Experimental NVFP4-quantized build of GLM 5.2 — offered for testing only. The published NVFP4 results look very flattering, but we're skeptical of its real quality and performance: this checkpoint was NOT QAT post-trained for NVFP4 (the QAT-on-NVFP4 models are where we've seen the best quality), so treat the benchmarks with caution. Play with it, push it, and see how far it gets you — for production work we recommend the fp8 `umans-glm-5.2`. It runs on a single low-capacity GPU with no fallback, so expect it to go down under load: crash it, let it restart, and play again.","weights":{"precision":"nvfp4","hf_url":"https://huggingface.co/nvidia/GLM-5.2-NVFP4"}},"umans-glm-5.1":{"display_name":"Umans GLM 5.1","provider":"GLM","retired_on":"2026-06-24","replacement":"umans-glm-5.2","note":"no longer available; use umans-glm-5.2","weights":{"precision":"fp8","hf_url":"https://huggingface.co/zai-org/GLM-5.1-FP8"}},"umans-kimi-k2.6":{"display_name":"Umans Kimi K2.6","provider":"Moonshot","retired_on":"2026-06-18","replacement":"umans-kimi-k2.7","note":"no longer available; use umans-kimi-k2.7","weights":{"precision":"full","hf_url":"https://huggingface.co/moonshotai/Kimi-K2.6"}},"umans-kimi-k2.5":{"display_name":"Umans Kimi K2.5","provider":"Moonshot","retired_on":"2026-05-13","replacement":"umans-kimi-k2.6","weights":{"precision":"full","hf_url":"https://huggingface.co/moonshotai/Kimi-K2.5"}},"umans-minimax-m2.5":{"display_name":"Umans MiniMax M2.5","provider":"MiniMax","retired_on":"2026-05-09","weights":{"hf_url":"https://huggingface.co/MiniMaxAI/MiniMax-M2.5"}}}