Baseten
AI model inference, training, model APIs, and self-hosted deployments.
Baseten is a GPU and AI hosting provider based in an undisclosed country. No European datacenter locations are recorded for it yet. A free tier is available. Its CLOUD Act exposure has not been verified. Pricing on this page is published only after a human has reviewed it.
Pricing
- BasicPay as you go. Includes dedicated deployments, model APIs, training, fast cold starts, SOC 2 Type II and HIPAA compliance, and email and in-app chat support.Monthly0 USDvCPUNot listedMemoryNot listedStorageNot listed
- ProVolume discounts available. Includes Basic features plus priority access to high-demand GPUs, dedicated compute, higher Model API rate limits, engineering expertise, and dedicated Slack and Zoom support.MonthlyNot listedvCPUNot listedMemoryNot listedStorageNot listed
- EnterpriseVolume discounts available. Includes custom SLAs, self-host deployments, on-demand flex compute, existing cloud commitments, data residency controls, advanced security and compliance, custom global regions, and advanced RBAC with Teams.MonthlyNot listedvCPUNot listedMemoryNot listedStorageNot listed
- Kimi K3Model API price per 1M tokens: Input $3.00, Cache Input $0.30, Output $15.00.Monthly$3.00vCPUNot listedMemoryNot listedStorageNot listed
- Kimi K2.6Model API price per 1M tokens: Input $0.95, Cache Input $0.16, Output $4.00.Monthly$0.95vCPUNot listedMemoryNot listedStorageNot listed
- Kimi K2.7 CodeModel API price per 1M tokens: Input $0.95, Cache Input $0.16, Output $4.00.Monthly$0.95vCPUNot listedMemoryNot listedStorageNot listed
- Inkling-SmallModel API price per 1M tokens: Input $0.50, Cache Input $0.10, Output $1.20.Monthly$0.50vCPUNot listedMemoryNot listedStorageNot listed
- InklingModel API price per 1M tokens: Input $1.00, Cache Input $0.17, Output $4.05.Monthly$1.00vCPUNot listedMemoryNot listedStorageNot listed
- GLM-5.2Model API price per 1M tokens: Input $1.40, Cache Input $0.14, Output $4.40.Monthly$1.40vCPUNot listedMemoryNot listedStorageNot listed
- GLM-5.2 FastModel API price per 1M tokens: Input $2.10, Cache Input $0.21, Output $6.60.Monthly$2.10vCPUNot listedMemoryNot listedStorageNot listed
- GLM 4.7Model API price per 1M tokens: Input $0.60, Cache Input $0.12, Output $2.20.Monthly$0.60vCPUNot listedMemoryNot listedStorageNot listed
- NVIDIA Nemotron 3 UltraModel API price per 1M tokens: Input $0.60, Cache Input $0.12, Output $2.40.Monthly$0.60vCPUNot listedMemoryNot listedStorageNot listed
- DeepSeek-V4-Flash-0731Model API price per 1M tokens: Input $0.13, Cache Input $0.028, Output $0.26.Monthly$0.13vCPUNot listedMemoryNot listedStorageNot listed
- DeepSeek V4 ProModel API price per 1M tokens: Input $1.74, Cache Input $0.145, Output $3.48.Monthly$1.74vCPUNot listedMemoryNot listedStorageNot listed
- GPT OSS 120BModel API price per 1M tokens: Input $0.10, Output $0.50. Cache Input is shown as '-'.Monthly$0.10vCPUNot listedMemoryNot listedStorageNot listed
- T4Dedicated Deployments and Training. Price per minute. 16 GiB VM.Monthly$0.01052vCPUNot listedMemoryNot listedStorageNot listed
- L4Dedicated Deployments and Training. Price per minute. 24 GiB VRAM.Monthly$0.01414vCPUNot listedMemoryNot listedStorageNot listed
- A10GDedicated Deployments and Training. Price per minute. 24 GiB VM.Monthly$0.02012vCPUNot listedMemoryNot listedStorageNot listed
- A100Dedicated Deployments and Training. Price per minute. 80 GiB VRAM.Monthly$0.06667vCPUNot listedMemoryNot listedStorageNot listed
- H100 MIGDedicated Deployments and Training. Price per minute. 40 GiB VRAM.Monthly$0.0625vCPUNot listedMemoryNot listedStorageNot listed
- H100Dedicated Deployments and Training. Price per minute. 80 GiB VRAM.Monthly$0.10833vCPUNot listedMemoryNot listedStorageNot listed
- B200Dedicated Deployments and Training. Price per minute. 180 GiB VRAM.Monthly$0.16633vCPUNot listedMemoryNot listedStorageNot listed
- 1x2Dedicated Deployments and Training. Price per minute.Monthly$0.00058vCPU1Memory2 GBStorageNot listed
- 1x4Dedicated Deployments and Training. Price per minute.Monthly$0.00086vCPU1Memory4 GBStorageNot listed
- 2x8Dedicated Deployments and Training. Price per minute.Monthly$0.00173vCPU2Memory8 GBStorageNot listed
- 4x16Dedicated Deployments and Training. Price per minute.Monthly$0.00346vCPU4Memory16 GBStorageNot listed
- 8x32Dedicated Deployments and Training. Price per minute.Monthly$0.00691vCPU8Memory32 GBStorageNot listed
- 16x64Dedicated Deployments and Training. Price per minute.Monthly$0.01382vCPU16Memory64 GBStorageNot listed
Pros and cons
What works
- Basic has a $0 monthly charge.
- Compute is billed only while models use compute.
- Self-hosted and hybrid deployments are available.
- SOC 2 Type II and HIPAA are stated.
What to watch
- Pro and Enterprise pricing is not published.
- Compute prices are usage based rather than fixed monthly allocations.
- The page does not list specific datacenter countries.
- Support level varies by plan.
Specifications
About Baseten
Baseten provides dedicated inference deployments, pre-optimized Model APIs, model training, and self-hosted or hybrid deployment options. Basic has no monthly platform charge and uses pay as you go pricing, while Pro and Enterprise pricing is quote based. Compute is metered by the minute, and Model APIs are priced per 1M tokens.
Reviews
No approved reviews yet. Every review is read by a person before it appears, and only approved reviews count toward a rating.
Other gpu and ai hosting
Something here out of date?
Every figure on this page was checked by a person, and providers change their plans without telling anybody. Say what moved and it gets checked again.
Work at Baseten? Claim this listing to correct the description and add screenshots. Verification is a confirmation link to an address on your own domain. Pricing still goes through the same check as every other listing here.