Kimi-K3-Uncensored-GGUF - 8× H100 80GB
kimi-k3-uncensored-gguf
kimi-k3-uncensored-gguf
About
Kimi-K3-Uncensored-GGUF is in the OpenLLM catalog on a dedicated 8× H100 80GB. Deploy it with a flat time pack and call your instance through an OpenAI-compatible API.
Compare
Model Cost Across Durations
Live pack pricing vs typical API estimates from 11 hours through 1 month.
Live pack pricing for this model — API competitor estimates coming soon.
Time pack
Kimi-K3-Uncensored-GGUF on 8× H100 80GB
24 hours cost
$805 USD
Lowest
Models in chart
- Kimi-K3-Uncensored-GGUF on 8× H100 80GB
At a glance
GPU
8× H100 80GB
GPUs
8
Memory
80GB
Apps & integrations
Choose an app below. Each guide shows how to point the app at your OpenAI-compatible endpoint.
n8n
Automate workflows and call your model as a node.
Open
OpenClaw
Build AI agents and tools on an OpenAI-compatible endpoint.
Open
Hermes
Connect agent runners to your chat completions endpoint.
Open
OpenCode
Power developer tools with your OpenAI-compatible model.
Open
Cursor
Override OpenAI Base URL in Cursor Settings and use your model with BYOK.
Open
VS Code
Use the Cline extension in VS Code to connect your OpenAI-compatible endpoint.
Open
Codex
Run OpenAI Codex CLI against your Chat Completions endpoint via config.toml.
Open
Raspberry Pi
Full Pi OS guide: SSH, API keys, curl, Python venv, systemd, and troubleshooting.
Open