Try VxCloud free for 7 days — no credit card required. Plans from $10/month. Get started →

VxCloud
The Most Advanced Multi-Cloud Platform

Multi-Cloud Infrastructure
Run by AI Agents

One control plane for AWS, Azure, Google Cloud, Oracle and Alibaba. Provision, deploy and optimise from a single workflow — with AI agents doing the repetitive work.

Agentic AI Automation
Multi-Cloud Management
Zero Downtime
Get Started Now

Multi-Cloud — Deploy Anywhere

AWSAWS
AzureAzure
GCPGCP
OracleOracle
AlibabaAlibaba
VultrVultr
TerraformTerraform
GitHubGitHub
8+ Integrations Active
99.99% uptime
Real screens

This is the actual product

Not an illustration. These are screenshots of the running dashboard — the same console you get on the free tier.

vxcloud.io/dashboard?tab=monitoring
VxCloud Monitoring — Every deployment, one control plane

Every deployment, one control plane

Live operations across AWS, Azure, GCP, Alibaba and bare metal — status, node, resource, session and provider in a single table, with per-resource detail pages you can link a colleague straight to.

Open this console free
AI-Native DevOps · Zero Downtime · Infinite Scale

Ship Faster with
Autonomous Cloud Ops

AI-driven bots that deploy, monitor, and self-heal your infrastructure around the clock — so your team can focus on building, not firefighting.

Infrastructure Agent

Automate cloud provisioning

Active

Database Agent

Manage & optimize databases

Active

Security Agent

Enforce security policies

Active

Monitoring Agent

Real-time observability

Active

Cost Agent

Optimize cloud spending

Active

Deployment Agent

Automate CI/CD pipelines

Active

Operations Agent

Manage day-to-day ops

Active

Multi-Cloud Agent

Orchestrate across clouds

Active

Optimization Agent

AI-powered insights

Active
Integrations & Open Source Stack

Integrate Seamlessly with the
AI & Cloud Tools You Rely On

Our platform connects natively with leading open-source AI frameworks, cloud providers, and developer tools — no glue code required.

LangChain
LangChainOSS
Anthropic Claude
Anthropic ClaudeAI
OpenAI
OpenAIAI
Groq
GroqAI
HuggingFace
HuggingFaceOSS
Mistral AI
Mistral AIAI
Meta / Llama
Meta / LlamaOSS
Microsoft
MicrosoftCloud
MCP Protocol
MCP ProtocolAI
Ollama
OllamaOSS
OpenClaw
OpenClawAgent
OpenCode
OpenCodeAI
AutoGen
AutoGenOSS
LangChain
LangChainOSS
Anthropic Claude
Anthropic ClaudeAI
OpenAI
OpenAIAI
Groq
GroqAI
HuggingFace
HuggingFaceOSS
Mistral AI
Mistral AIAI
Meta / Llama
Meta / LlamaOSS
Microsoft
MicrosoftCloud
MCP Protocol
MCP ProtocolAI
Ollama
OllamaOSS
OpenClaw
OpenClawAgent
OpenCode
OpenCodeAI
AutoGen
AutoGenOSS
AWS
AWSCloud
Azure
AzureCloud
Google Cloud
Google CloudCloud
Oracle
OracleCloud
Alibaba Cloud
Alibaba CloudCloud
Vultr
VultrCloud
Kubernetes
KubernetesOSS
Docker
DockerOSS
Red Hat
Red HatOSS
Terraform
TerraformIaC
Ansible
AnsibleIaC
GitHub Actions
GitHub ActionsCI/CD
Prometheus
PrometheusOSS
Grafana
GrafanaOSS
OpenTelemetry
OpenTelemetryOSS
Datadog
DatadogObs
Elastic
ElasticOSS
Vault
VaultSec
Dynatrace
DynatraceObs
AWS
AWSCloud
Azure
AzureCloud
Google Cloud
Google CloudCloud
Oracle
OracleCloud
Alibaba Cloud
Alibaba CloudCloud
Vultr
VultrCloud
Kubernetes
KubernetesOSS
Docker
DockerOSS
Red Hat
Red HatOSS
Terraform
TerraformIaC
Ansible
AnsibleIaC
GitHub Actions
GitHub ActionsCI/CD
Prometheus
PrometheusOSS
Grafana
GrafanaOSS
OpenTelemetry
OpenTelemetryOSS
Datadog
DatadogObs
Elastic
ElasticOSS
Vault
VaultSec
Dynatrace
DynatraceObs

& many more — built on open standards, not vendor lock-in.

120+
Teams Using VxCloud
8,500+
Deployments Automated
30+
Cloud Regions Supported
99.99%
Uptime SLA

Platform Pillars

Multi-cloud deployment & management, Observability, Cost, Security

Deploy across AWS, Azure, Google Cloud, Oracle, Alibaba with unified workflows. Gain deep visibility, optimize cloud spend, and enforce enterprise-grade security and compliance.

Multi-cloud deployment

One pipeline to deploy to AWS, Azure, Google Cloud, Oracle, Alibaba and on-prem. Native IaC and GitOps.

Observability

Integrated logs, metrics, traces with alerting. SLO dashboards out of the box.

Cost optimization

Ownership, showback, and recommendations to reduce spend without sacrificing SLAs.

Security & compliance

Policy-as-code, least privilege by default, SOC 2 / ISO 27001 ready.

Features

Everything you need to build

Infrastructure as Code

Define your stack in Terraform, Pulumi, or our visual builder. Every change is versioned, reviewed, and deployed through automated pipelines.

Multi-Region Deployments

Deploy to 30+ cloud regions across AWS, Azure, Google Cloud, Oracle, and Alibaba with built-in traffic routing, failover, and edge caching for global delivery.

Zero Trust Security

Secrets management, network policy enforcement, RBAC, and automated vulnerability scanning baked into every deployment pipeline.

Solutions

Cloud Solutions

Cloud Computing

Auto-scaling compute across providers — VMs, containers, and serverless with unified cost visibility

Data Infrastructure

Managed databases, streaming pipelines, and data warehouses with automated backups and failover

AI/ML Workloads

GPU provisioning, model serving, and experiment tracking integrated into your deployment pipeline

Infrastructure-as-code, in your language

Built for developers & enterprises — no proprietary DSL.

Provision and deploy vxcloud services with our SDKs (Go · Python · TypeScript), the CLI, or a GitHub Action — a few lines, every cloud.

provision.py live
import vxsdk
 
c = vxsdk.Client.load_from_vxcli()
 
# Managed Redis — a few lines, any cloud
db = c.cloud.create_redis(
"my-redis", cloud="aws", region="us-east-1",
node_type="cache.t3.micro",
)
print(db["endpoint"])

$ pip install vxsdk && python provision.py

✓ authenticated → node1.vxcloud.io

✓ redis "my-redis" provisioned · us-east-1

→ my-redis.cache.vxcloud.io:6379

CI/CD Pipelines

Enable CI/CD

on your projects

Use our fully managed CI/CD system to deploy your applications and backend services to your favorite cloud provider — automatically, on every push.

01Push Code
02Run Tests
03Build Image
🚀
04Deploy Live

Works with your stack

+ 10 more — Flask · Laravel · Spring Boot · Rust · Java · C++ · PHP · Streamlit · Expo · Static

Pick your framework — VxCloud generates the GitHub Actions workflow and runs the matching vxcli deploy over SSH on every push. No DSL, no Terraform provider.

Live pipeline example

deploy-nextjs.ymlRunning
name: Deploy Next.js to vxcloud
on:
push:
branches: [main]
jobs:
deploy:
runs-on: ubuntu-latest
steps:
- uses: prodxcloud/vxcloud-deploy-action@v1
with:
username: ${{ secrets.VXCLOUD_USERNAME }}
api-key: ${{ secrets.VXCLOUD_API_KEY }}
command: deploy nextjs --app-name web --build-mode production
--host ${{ vars.HOST }} --ssh-user ubuntu --key-pair-name VPS1

$ git push origin main

→ Installing vxcli… ✓ v2026.8.14 (checksum verified)

✓ Authenticated as joelwembo → node1.vxcloud.io

🚀 vxcli deploy nextjs --app-name web --build-mode production

· building .next/ · SSR + standalone output

✓ Deployment initiated — exit_code 0

🌐 Next.js live at https://web.vxcloud.online

Integrations

Works with your favorite tools

GitHub

GitHub

Connect your repositories and automate deployments

Docker

Docker

Deploy containerized applications

Kubernetes

Kubernetes

Orchestrate containerized applications

Automate Your DevOps with AI

Describe your infrastructure needs and let AI generate production-ready configurations

VMs
Cloud
Databases
CI/CD
Deploy
GitOps

Latest Writing

Latest from VxCloud & Joel Wembo Technologies

What's happening across cloud, AI, DevOps, product launches, and platform engineering right now.

6 latest storiesExplore all articles

Compare

Multi-Cloud Cost Calculator

Estimate monthly spend across 12 cloud providers side-by-side. List prices, no API calls.

Multi-Cloud Cost Calculator

List prices as of 2026-04 · 12 providers · no API calls

24/7 always-on

Approx. $0.10/GB-month SSD across clouds

Approx. $0.09/GB after free tier

Amazon Web Services

$19.00/mo

no plan

Microsoft Azure

$19.00/mo

no plan

Google Cloud Platform

$19.00/mo

no plan

Public on-demand list prices, Linux, no commitments. GPU prices include host compute. Regional prices shown reflect us-east-1 / eastus / us-central1 or closest equivalent. Hetzner/UpCloud bill in EUR with monthly caps; hourly is approximated at ~1.08 FX. Alibaba prices vary 3-8x by region — values shown are Singapore/SG where available. Confirmed against aws-pricing.com, instances.vantage.sh, cloudprice.net, gcloud-compute.com, computeprices.com, sparecores.com, costgoat.com, getdeploying.com, digitalocean.com/pricing, linode.com/pricing. Always verify against the provider's official calculator before procurement.

Benchmark

AI Model Pricing

Compare input/output token prices, context windows, and benchmark scores across leading AI models.

AI Model Pricing

241 models · 34 creators · list prices per 1M tokens, as of 2026-04 · no API calls

Prefix with / for regex, e.g. /^GPT-5/

$0$150
$0$600
Showing 241 of 241 modelsPrices per 1M tokens · USD
ModelCreatorInput $Output $ProvidersContextMax OutputIntelligenceCoding
GPT-5.5OpenAI5.0030.001272K128K60.259.1
Claude Opus 4.7Anthropic5.0025.0071M128K57.352.5
Gemini 3.1 Pro PreviewGoogle2.0012.0041M66K57.255.5
GPT-5.4OpenAI2.5015.0051.1M128K56.857.3
Kimi K2.6Moonshot AI (Kimi)0.7454.004262K66K53.947.1
MiMo V2.5 ProXiaomi1.003.0011M131K53.845.5
GPT-5.3 CodexOpenAI1.7514.004400K128K53.653.1
Qwen3.6 Max PreviewAlibaba1.307.801240K64K51.844.9
GLM-5.1Zhipu AI1.053.503203K66K51.443.4
GPT-5.2OpenAI1.7514.006410K128K51.348.7
Qwen3.6 PlusAlibaba0.3251.9521M66K50.042.9
GLM-5Zhipu AI0.5732.089203K131K49.844.2
MiniMax M2.7MiniMax0.3001.203205K131K49.641.9
MiMo V2 ProXiaomi1.003.0021M131K49.241.4
GPT-5.2 CodexOpenAI1.7514.004400K128K49.043.0
GPT-5.4 MiniOpenAI0.7504.5041.1M128K48.951.5
Grok 4.20xAI2.006.0032MN/A48.540.5
Gemini 3 ProGoogle2.0012.003N/AN/A48.446.5
GPT-5.1OpenAI1.2510.006410K128K47.744.7
Kimi K2.5Moonshot AI (Kimi)0.4402.0011262K98K46.839.5
GLM-5 TurboZhipu AI1.204.002203K131K46.836.8
DeepSeek V4 FlashDeepSeek0.1400.28031M384K46.538.7
Claude Opus 4.6Anthropic5.0025.0071M128K46.547.6
Qwen3.5 397B A17BAlibaba0.3902.344262K66K45.041.3
GPT-5 CodexOpenAI1.2510.004400K128K44.638.9
GPT-5OpenAI1.2510.008410K128K44.636.0
Claude Sonnet 4.6Anthropic3.0015.0071M128K44.446.4
GPT-5.4 NanoOpenAI0.2001.2541.1M128K44.043.9
KwaiPilot KAT 2 Pro CoderKwaiPilot0.3001.202256K80K43.845.6
MiMo V2 OmniXiaomi0.4002.001262K66K43.435.5
GPT-5.1 CodexOpenAI1.2510.004400K128K43.136.6
Claude Opus 4.5Anthropic5.0025.009410K64K43.142.9
GLM-5V TurboZhipu AI1.204.002203K131K42.936.2
Qwen3.5 27BAlibaba0.1951.562262K66K42.134.9
GLM-4.7Zhipu AI0.3801.7413205K131K42.136.3
MiniMax M2.5MiniMax0.1501.1571M131K41.937.4
Qwen3.5 122B A10BAlibaba0.2602.083262K66K41.634.7
Grok 4xAI3.0015.006256KN/A41.540.5
GPT-5 MiniOpenAI0.2502.007400K128K41.235.3
Kimi K2 ThinkingMoonshot AI (Kimi)0.5741.2013262K33K40.934.8
o3 ProOpenAI20.0080.004200K100K40.7N/A
Qwen3 Max ThinkingAlibaba0.7803.902262K66K39.930.5
MiniMax M2.1MiniMax0.2900.95081M131K39.432.8
Grok 4.1 ReasoningxAI0.2000.50032M30K38.630.9
GPT-5.1 Codex MiniOpenAI0.2502.004400K128K38.636.4
o3OpenAI2.008.005200K100K38.438.4
Step 3.5 FlashStepFun0.1000.3001262K66K37.831.6
Qwen3.5 35B A3BAlibaba0.1631.302262K66K37.130.3
Claude Sonnet 4.5Anthropic3.0015.00101M64K37.133.5
MiniMax M2MiniMax0.2551.009205K131K36.129.2
Nemotron Super 3 120B A12BNVIDIA0.0900.4502262K32K36.031.2
KwaiPilot KAT 1 Pro CoderKwaiPilot0.0301.201256K32K36.018.3
Claude Opus 4.1Anthropic15.0075.007200K32K36.0N/A
Grok 4 ReasoningxAI0.2000.50032M256K35.127.4
Gemini 3 FlashGoogle0.5003.0031M65K35.037.8
Claude Sonnet 3.7Anthropic3.0015.001200K64K34.727.6
Gemini 3.1 Flash Lite PreviewGoogle0.2501.5041M66K33.530.1
GPT OSS 120BOpenAI0.0390.19020131K131K33.328.6
o4 MiniOpenAI1.004.006200K100K33.125.6
Claude Sonnet 4Anthropic3.0015.00101M64K33.030.6
Claude Opus 4Anthropic15.0075.009410K32K33.0N/A
Mercury 2Inception0.2500.7502128K50K32.830.6
Qwen3.5 9BAlibaba0.1000.1502262KN/A32.425.3
DeepSeek V3.2DeepSeek0.2520.37812164K66K32.134.6
Arcee Trinity Large ThinkingArcee AI0.2200.8502262K80K31.927.2
Qwen3 MaxAlibaba0.3591.434262K66K31.426.4
Claude Haiku 4.5Anthropic1.005.009200K64K31.129.6
o1OpenAI15.0060.005200K100K30.820.5
Claude Sonnet 3.7 (200K)Anthropic3.0015.0010200K128K30.826.7
MiMo V2 FlashXiaomi0.0900.2904262K66K30.425.8
GLM-4.6Zhipu AI0.3901.9010205K131K30.230.2
GLM-4.7 FlashZhipu AI0.0600.4003203K131K30.125.9
Gemini 2.5 ProGoogle1.2510.0071M66K29.531.9
DeepSeek V3.2 SpecialeDeepSeek0.4001.202164KN/A29.437.9
Grok 4.20 Non-ReasoningxAI2.006.0022MN/A29.022.0
Grok 1 Code FastxAI0.2001.504256K10K28.723.7
DeepSeek V3.1 TerminusDeepSeek0.2100.7906164K66K28.531.9
Qwen3 Coder NextAlibaba0.1500.8005262K66K28.322.9
DeepSeek V3.1DeepSeek0.1350.50014164K66K28.128.4
GPT-5 NanoOpenAI0.0500.4007400K128K26.820.3
GLM-4.5Zhipu AI0.4001.608131K98K26.426.3
Kimi K2Moonshot AI (Kimi)0.4002.004262K128K26.322.1
GPT-4.1OpenAI2.008.0061M33K26.321.8
Qwen3 Max PreviewAlibaba1.206.002262K66K26.125.5
Solar Pro 3Upstage0.1500.6001128KN/A25.913.3
o3 MiniOpenAI1.104.405200K100K25.917.9
o1 ProOpenAI150.00600.003200K100K25.8N/A
Gemini 2.5 Flash PreviewGoogle0.3002.5021M66K25.722.1
o3 Mini HighOpenAI1.104.401200K100K25.217.3
Grok 3xAI3.0015.005131KN/A25.219.8
Qwen3 Coder 480B A35B InstructAlibaba0.2201.308262K66K24.824.6
Sonar Reasoning ProPerplexity2.008.003128K8K24.6N/A
GPT OSS 20BOpenAI0.0300.14016131K131K24.518.5
MiniMax M1 80kMiniMax0.1000.10031M40K24.414.5
o1 PreviewOpenAI15.0060.002128K33K23.734.0
Grok 4.1 FastxAI0.2000.50032M30K23.619.5
GLM-4.5 AirZhipu AI0.1300.8506131K98K23.223.8
Grok 4 FastxAI0.2000.50042M30K23.119.0
GPT-4.1 MiniOpenAI0.4001.6051M33K22.918.5
Mistral Large 3Mistral AI0.5001.204262K8K22.822.7
DeepSeek V3 324DeepSeek0.2000.40012164K16K22.322.0
INTELLECT-3Prime Intellect0.2001.102131KN/A22.219.1
Devstral 2Mistral AI0.4002.001256KN/A22.023.7
Mistral Medium 3.1Mistral AI0.4002.002131KN/A21.318.3
Qwen3 VL 235B A22B InstructAlibaba0.2000.8806262K33K20.816.5
o1 MiniOpenAI1.104.403128K66K20.4N/A
Qwen3 Next 80B A3B InstructAlibaba0.0900.90010262K66K20.115.3
Qwen3 Coder 30B A3B InstructAlibaba0.0700.2605262K66K20.019.4
GPT-4.5OpenAI75.00150.001N/AN/A20.0N/A
QwQ 32BAlibaba0.1500.2009131K16K19.7N/A
Devstral Small 2Mistral AI0.1000.3001256KN/A19.520.7
Gemini 2.5 Flash Lite PreviewGoogle0.1000.40031M66K19.414.5
Nova PremierAmazon2.5012.5031M32K19.013.8
Mistral Medium 3Mistral AI0.4002.002131K8K18.813.6
Magistral MediumMistral AI2.005.002128K64K18.816.0
DeepSeek R1DeepSeek0.2800.40015164K66K18.815.9
Devstral MediumMistral AI0.4002.002256KN/A18.715.9
Claude Haiku 3.5Anthropic0.8004.007200K8K18.710.7
Gemini 2.0 FlashGoogle0.1000.40051M8K18.513.6
Llama 4 MaverickMeta0.1500.60061M16K18.415.6
Nova 2 LiteAmazon0.3002.5031M66K18.012.5
Claude Opus 3Anthropic15.0075.005200K4K18.019.5
Sonar ReasoningPerplexity1.005.002128K8K17.9N/A
Gemini 2.5 FlashGoogle0.1500.60091M66K17.817.8
Llama 3.1 405B InstructMeta0.1200.30011131K16K17.414.5
Qwen3 VL 32B InstructAlibaba0.1040.4163131K33K17.215.6
DeepSeek R1 Distill Qwen 32BDeepSeek0.1500.1507131K32K17.2N/A
GLM-4.6VZhipu AI0.3000.9004131K33K17.111.1
Qwen3 235B A22B InstructAlibaba0.0900.5809262K33K17.014.0
Magistral SmallMistral AI0.5001.503128K64K16.811.1
DeepSeek V3DeepSeek0.2000.20012164K82K16.516.4
Qwen3 VL 30B A3B InstructAlibaba0.1300.5205262K33K16.114.3
Ministral 3 14BMistral AI0.2000.2001262KN/A16.010.9
DeepSeek R1 Distill Llama 70BDeepSeek0.2000.37511131K16K16.011.4
DeepSeek R1 Distill Qwen 14BDeepSeek0.0700.0704131K16K15.8N/A
Qwen2.5 72B InstructAlibaba0.1200.3008131K16K15.611.9
SonarPerplexity1.001.003128K8K15.5N/A
Sonar ProPerplexity3.0015.003200K8K15.2N/A
QwQ 32B PreviewAlibaba0.2870.861233K16K15.2N/A
Devstral SmallMistral AI0.1000.3004256K64K15.212.1
Mistral Small 3.2Mistral AI0.0600.1801131KN/A15.113.3
Mistral Large 2Mistral AI2.006.001128K8K15.113.8
Qwen3 30B A3BAlibaba0.0800.2807131K20K15.014.2
ERNIE 4.5 300B A47BBaidu0.2801.101123K12K15.014.5
Ministral 3 8BMistral AI0.1500.1501262KN/A14.810.0
Gemini 2.0 Flash LiteGoogle0.0750.30041M8K14.7N/A
Llama 3.3 70B InstructMeta0.1000.20019131K120K14.510.7
GPT-4oOpenAI2.5010.006131K16K14.516.6
Qwen3 VL 8B InstructAlibaba0.0800.2005131K33K14.37.3
Claude Sonnet 3.5Anthropic3.0015.0061M8K14.226.0
Pixtral Large (25.02)Mistral AI2.006.005131K4K14.0N/A
Grok 2xAI2.0010.002131K4K13.9N/A
GPT-4 TurboOpenAI10.0030.004128K4K13.721.5
Nova ProAmazon0.8003.204300K10K13.511.0
Llama 4 ScoutMeta0.0800.3005328K16K13.56.7
Command ACohere1.561.565256K8K13.59.9
Llama 3.1 Nemotron 70B InstructNVIDIA0.6000.6003131K16K13.410.8
Grok BetaxAI5.0015.001131KN/A13.3N/A
Qwen2.5 32B InstructAlibaba0.0600.2003131K8K13.2N/A
Nemotron Nano 3 30B A3BNVIDIA0.0500.2002262KN/A13.215.8
Nemotron Nano 2 9BNVIDIA0.0400.1605131K8K13.27.5
GPT-4.1 NanoOpenAI0.1000.40051M33K13.011.2
Qwen2.5 Coder 32B InstructAlibaba0.0500.1008131K8K12.9N/A
GPT-4OpenAI30.0060.00333K8K12.813.1
Nova LiteAmazon0.0600.2404300K10K12.75.1
GLM-4.5VZhipu AI0.6001.206131K32K12.710.8
Gemini 2.5 Flash LiteGoogle0.0750.30061M66K12.77.4
GPT-4o miniOpenAI0.1500.6006131K16K12.6N/A
Qwen3 4B InstructAlibaba0.0100.0302262K33K12.59.1
Qwen3 30B A3B InstructAlibaba0.0900.3003262K33K12.513.3
Llama 3.1 70B InstructMeta0.1000.10013131K8K12.510.9
DeepSeek V2.5DeepSeek1.201.20133KN/A12.5N/A
Claude Haiku 3Anthropic0.2501.255200K4K12.36.7
OLMo 3.1 32B InstructAllen AI0.2000.600166KN/A12.25.6
OLMo 3 32B ThinkAllen AI0.1500.500166K4K12.110.5
Mistral SabaMistral AI0.2000.600133KN/A12.1N/A
DeepSeek R1 Distill Llama 8BDeepSeek0.0250.0253131KN/A12.1N/A
Reka FlashReka AI0.9000.9001100K8K12.0N/A
Qwen TurboAlibaba0.0330.13021M16K12.0N/A
Gemini 1.5 ProGoogle1.255.002N/AN/A12.019.8
Llama 3.1 8B InstructMeta0.0200.03019200K128K11.84.9
Qwen2 72B InstructAlibaba0.9000.900133KN/A11.7N/A
Ministral 3 3BMistral AI0.1000.1001131KN/A11.24.8
Gemini 1.5 Flash 8BGoogle0.0370.3002N/AN/A11.1N/A
Jamba 1.7 LargeAI21 Labs2.008.002256K4K10.97.8
Granite 4 Small HIBM0.0600.250120KN/A10.88.5
Qwen3 Omni 30B A3B InstructAlibaba0.2500.970166K16K10.77.2
Jamba 1.5 LargeAI21 Labs2.002.804256K8K10.7N/A
Jamba 1.6 LargeAI21 Labs2.008.001256KN/A10.6N/A
Hermes 3 Llama 3.1 70BNous Research0.1200.3003131KN/A10.6N/A
LFM 2 24B A2BLiquid AI0.0300.120133KN/A10.53.6
Gemini 1.5 FlashGoogle0.0750.30028KN/A10.5N/A
Phi-4Microsoft0.0650.140316K16K10.411.2
Nova MicroAmazon0.0350.1404128K10K10.34.1
Claude Sonnet 3Anthropic3.0015.002200K4K10.3N/A
Nemotron Nano 2 12B VLNVIDIA0.1000.1003131K4K10.15.9
Qwen2.5 Coder 7B InstructAlibaba0.0100.0304131K8K10.0N/A
Mistral LargeMistral AI0.5001.508262K16K9.9N/A
Mixtral 8x22B InstructMistral AI0.6000.600666K2K9.8N/A
Llama 3.2 3B InstructMeta0.0150.0209131K32K9.7N/A
Llama 2 7B ChatMeta0.0500.15044K4K9.7N/A
Reka Flash 3Reka AI0.1000.200166KN/A9.58.9
DeepSeek R1 Distill Qwen 1.5BDeepSeek0.0900.0903131KN/A9.1N/A
Claude 2Anthropic8.0024.001100K8K9.112.9
Mistral SmallMistral AI0.0600.1807262K8K9.0N/A
Mistral MediumMistral AI0.4002.004131K64K9.0N/A
GPT-3.5 TurboOpenAI0.5001.50416K4K9.010.7
Llama 3 70B InstructMeta0.1200.3009131K8K8.96.8
LFM 40BLiquid AI0.1000.2001131KN/A8.8N/A
Gemma 3 12BGoogle0.1500.5001128K32K8.86.3
Gemini 1.0 ProGoogle0.5001.502N/AN/A8.5N/A
Phi-4 MiniMicrosoftN/AN/A1N/AN/A8.43.6
Llama 2 70B ChatMeta0.5000.90074K4K8.4N/A
Llama 2 13B ChatMeta0.1000.20044K4K8.4N/A
Command R+Cohere1.561.566128K4K8.3N/A
Jamba 1.7 MiniAI21 Labs0.2000.4001256KN/A8.13.1
Jamba 1.5 MiniAI21 Labs0.2000.2004256K8K8.0N/A
Jamba 1.6 MiniAI21 Labs0.2000.4001256KN/A7.9N/A
Mixtral 8x7B InstructMistral AI0.0700.1501033K16K7.7N/A
Mistral 7B InstructMistral AI0.0100.1009127K16K7.4N/A
Command RCohere0.1500.1505128K4K7.4N/A
Granite 3.3 8B InstructIBM0.0300.20028K8K7.03.4
Llama 3 8B InstructMeta0.0300.040932K8K6.44.0
Llama 3.2 1B InstructMeta0.0270.0805128K16K6.30.6
Gemma 3 4BGoogle0.0300.0801128K8K6.32.9
AdaOpenAI0.100N/A18KN/AN/AN/A
Aion 1Aion Labs4.008.001131K33KN/AN/A
Aion 1 MiniAion Labs0.7001.401131K33KN/AN/A
Aion 2Aion Labs0.8001.601131K33KN/AN/A
Aion Llama 3.1 RP 8BAion Labs0.8001.60133KN/AN/AN/A
Allam 1 13B InstructSDAIA1.801.8018K8KN/AN/A
Apply 3Relace0.8501.251256K128KN/AN/A
Arcee Coder LargeArcee AI0.5000.800133KN/AN/AN/A
Arcee Maestro ReasoningArcee AI0.9003.301131K32KN/AN/A
Arcee SpotlightArcee AI0.1800.1801131K66KN/AN/A
Arcee Trinity Large PreviewArcee AI0.1500.4502131KN/AN/AN/A
Arcee Trinity MiniArcee AI0.0450.1502131KN/AN/AN/A
Arcee Virtuoso LargeArcee AI0.7501.201131K64KN/AN/A
ArcticSnowflake1.681.6814KN/AN/AN/A
Arctic 1.5 Embed MSnowflake0.0600.0601N/AN/AN/AN/A
Arctic 2 Embed LSnowflake0.1000.1001N/AN/AN/AN/A

Public list prices per 1M tokens (input/output) for LLM API inference. Intelligence & Coding scores reflect public benchmarks; higher is better. 'providers' is the count of inference providers offering the model. Prices vary over time and by region/tier — verify against the provider's pricing page before procurement.

Watch

Platform demos & industry talks

Learn how teams ship faster on multi-cloud and build robust observability practices.

Explore VxCloud for enterprises →

Trusted by leading companies worldwide

AWS
Azure
Google
IBM
Meta
Intel
Apple
NVIDIA
AWS
Azure
Google
IBM
Meta
Intel
Apple
NVIDIA
AI Robot
Google Logo

Built for AI-scale infrastructure

Power demanding AI workloads with confidence

What our customers say

A

Alex Kimura

Platform Engineer at NovaBridge

"VxCloud completely streamlined our release workflow. We went from manual weekend deployments to fully automated pipelines shipping multiple times a day."

R

Rachel Okonkwo

CTO at Luminary Labs

"The managed infrastructure on VxCloud just works. We scaled from a small prototype to handling millions of requests without ever worrying about uptime."

J

James Petrov

Head of Security at Crestfield IO

"Security was non-negotiable for us. VxCloud gave us enterprise-grade compliance controls out of the box, so our team could focus on building instead of auditing."

Frequently Asked Questions

Ready to modernize your cloud operations?Register now and deploy with confidence on professional-grade infrastructure.