Deltafin Cloud
Private LLM hosting with Kimi‑K3
Deltafin Cloud offers a self‑hosted, OpenAI‑compatible API that runs the full Kimi‑K3 model on a single GPU server. Clients can deploy the service in their own data center or on a cloud VM, keeping data private while enjoying low‑latency inference.
FastAPI + Docker + PyTorch + CUDA + Hugging Face Hub + Stripe + Vercel
Not required
Free tier: 100k tokens/month, 1 GPU, 1 concurrent request. Pro tier: unlimited tokens, 2 GPUs, 5 concurrent requests, priority support
Medium — Requires a GPU, 1.7 TB of storage, and native library builds, but can be packaged into a Docker image for repeatable deployment
- ✓Large star count indicates community interest
- ✓MIT license allows commercial use
- ✓OpenAI‑compatible API simplifies integration for existing clients
Related analyses
ChatHistory Hub
Centralize your AI coding assistant logs
Lunel Cloud
Deploy and manage secure proxy instances in minutes
TokenTab Pro
Offline AI cost tracking for teams