← Portal
/
🖥️ GPU Sidecar
zephai-gpu
Standard_NC4as_T4_v3 · NVIDIA T4 16 GB VRAM · West US 2
Checking...
Compute cost
—
LLM provider
—
◌
VM booting
— Azure allocating T4 hardware
~2 min
◌
vLLM loading model
— transferring weights to VRAM
~2 min
KV Cache
—
Active Requests
—
Prefix Cache Hit
—
▶ Start GPU VM
⚡ Switch to vLLM
↩ Switch to Bedrock
■ Stop VM
⚠️ Stopping will automatically switch the LLM provider back to Bedrock before shutting down the VM.
VM DETAILS
Resource Group
zephai-prod-rg
VM Size
Standard_NC4as_T4_v3
Private IP
10.0.0.10
vLLM endpoint
http://10.0.0.10:8001/v1
Public IP
None (internal only)
Auto-shutdown
00:00 UTC daily
On-demand cost
$0.53/hr compute + ~$0.03/hr disk
Connecting...
🖥️
Infrastructure
Azure resources · auto-refreshes every 30s
⚡
App VM —
zephai-vm
Hosts ZephAI API, worker, nginx, Redis, ChromaDB
ℹ️
Stop VM
🐘
PostgreSQL —
zephai-postgres
Flexible server · stop to pause compute billing
ℹ️
…
🖥️
Details
✕
Loading…