Liquid AI Infra & Persistent GPU Cloud — A100 / H100 / B200 / B300 / GB200
On the Hub: vessl/Kimi-K3-W4AFP8 — Kimi K3 quantized to W4A-FP8 for efficient sglang inference.
sglang
🌐 vessl.ai · 💻 GitHub · 📝 Blog · 🐦 X · 💼 LinkedIn · ▶️ YouTube