Rent A100 40GB GPU capacity

Run models, scripts, containers, training jobs, batch inference, and custom GPU workloads on A100 40GB-class capacity with max-cost protection.

Checking last-seen pricing...

Popular A100 40GB use cases

fine-tuningLoRA adapterscheckpoints7B trainingcustom jobs

Model training, fine-tuning, LoRA adaptation

Models people serve on A100 40GB

Qwen 2.5 7B InstructQwen 2.5 7B Instruct AWQQwen 2.5 7B Instruct GPTQQwen3 8BQwen3 8B AWQQwen3 8B GPTQLlama 3.1 8B InstructLlama 3.1 8B Instruct AWQLlama 3.1 8B Instruct GPTQLlama 3.2 11B Vision Instruct AWQLlama 3.2 11B Vision Instruct GPTQDeepSeek R1 Distill Qwen 7BDeepSeek R1 Distill Qwen 7B AWQDeepSeek R1 Distill Qwen 7B GPTQGemma 3 4B ITGemma 3 4B IT AWQGemma 3 4B IT GPTQGemma 3 12B IT AWQGemma 3 12B IT GPTQGemma 3 27B IT AWQGemma 3 27B IT GPTQMistral 7B Instruct v0.3Mistral 7B Instruct v0.3 AWQMistral 7B Instruct v0.3 GPTQMistral Small 24B Instruct AWQMistral Small 24B Instruct GPTQPhi 3.5 Mini InstructPhi 3.5 Mini Instruct AWQPhi 3.5 Mini Instruct GPTQPhi-4 AWQPhi-4 GPTQCommand R7BCommand R7B AWQCommand R7B GPTQGranite 3.3 8B InstructGranite 3.3 8B Instruct AWQGranite 3.3 8B Instruct GPTQNemotron Mini 4B InstructNemotron Mini 4B Instruct AWQNemotron Mini 4B Instruct GPTQGPT-OSS 20B AWQGPT-OSS 20B GPTQBGE-M3BGE-M3 AWQBGE-M3 GPTQBGE Reranker v2 M3BGE Reranker v2 M3 AWQBGE Reranker v2 M3 GPTQNomic Embed Text v1.5Nomic Embed Text v1.5 AWQNomic Embed Text v1.5 GPTQWhisper Large v3Whisper Large v3 AWQWhisper Large v3 GPTQLLaVA OneVision Qwen2 7BLLaVA OneVision Qwen2 7B AWQLLaVA OneVision Qwen2 7B GPTQFalcon3 10B InstructFalcon3 10B Instruct AWQFalcon3 10B Instruct GPTQYi 1.5 9B ChatYi 1.5 9B Chat AWQYi 1.5 9B Chat GPTQStarCoder2 15B AWQStarCoder2 15B GPTQCodestral 22B AWQCodestral 22B GPTQJina Embeddings v3Jina Embeddings v3 AWQJina Embeddings v3 GPTQMultilingual E5 LargeMultilingual E5 Large AWQMultilingual E5 Large GPTQGTE Large EN v1.5GTE Large EN v1.5 AWQGTE Large EN v1.5 GPTQModernBERT LargeModernBERT Large AWQModernBERT Large GPTQOpenHermes 2.5 Mistral 7BOpenHermes 2.5 Mistral 7B AWQOpenHermes 2.5 Mistral 7B GPTQOpenChat 3.5 0106OpenChat 3.5 0106 AWQOpenChat 3.5 0106 GPTQStable Code Instruct 3BStable Code Instruct 3B AWQStable Code Instruct 3B GPTQMiniCPM5 1BMiniCPM5 1B AWQMiniCPM5 1B GPTQOLMo 2 0425 1B InstructOLMo 2 0425 1B Instruct AWQOLMo 2 0425 1B Instruct GPTQLFM2.5 1.2B InstructLFM2.5 1.2B Instruct AWQLFM2.5 1.2B Instruct GPTQQwen 2.5 Coder 3B InstructQwen 2.5 Coder 3B Instruct AWQQwen 2.5 Coder 3B Instruct GPTQQuokka 7B InstructQuokka 7B Instruct AWQQuokka 7B Instruct GPTQEXAONE 3.5 7.8B InstructEXAONE 3.5 7.8B Instruct AWQEXAONE 3.5 7.8B Instruct GPTQA.X 4.0 LightA.X 4.0 Light AWQA.X 4.0 Light GPTQStarCoder2 3B InstructStarCoder2 3B Instruct AWQStarCoder2 3B Instruct GPTQKanana 2 3B InstructKanana 2 3B Instruct AWQKanana 2 3B Instruct GPTQOlmo 3 7B InstructOlmo 3 7B Instruct AWQOlmo 3 7B Instruct GPTQQwen 2.5 1.5B InstructQwen 2.5 1.5B Instruct AWQQwen 2.5 1.5B Instruct GPTQQwen 2.5 1.5B Instruct (Gensyn mirror)Qwen 2.5 1.5B Instruct (Gensyn mirror) AWQQwen 2.5 1.5B Instruct (Gensyn mirror) GPTQQwen 2.5 0.5B InstructQwen 2.5 0.5B Instruct AWQQwen 2.5 0.5B Instruct GPTQFalcon-H1 0.5B InstructFalcon-H1 0.5B Instruct AWQFalcon-H1 0.5B Instruct GPTQQwen 3 0.6BQwen 3 0.6B AWQQwen 3 0.6B GPTQQwen 3 0.6B FP8Qwen 3 0.6B FP8 AWQQwen 3 0.6B FP8 GPTQGemma 4 E4B Instruct GGUFGemma 4 E4B Instruct GGUF AWQGemma 4 E4B Instruct GGUF GPTQQwen 3.5 4B MTP GGUFQwen 3.5 4B MTP GGUF AWQQwen 3.5 4B MTP GGUF GPTQStartLux Decision 0.8B Q8_0 GGUFStartLux Decision 0.8B Q8_0 GGUF AWQStartLux Decision 0.8B Q8_0 GGUF GPTQGemma 4 E2BGemma 4 E2B AWQGemma 4 E2B GPTQQwen 3.8 27B GGUF AWQQwen 3.8 27B GGUF GPTQMiniCPM5 2BMiniCPM5 2B AWQMiniCPM5 2B GPTQEmbeddingGemma 2EmbeddingGemma 2 AWQEmbeddingGemma 2 GPTQQwen-Image 2.1 TurboQwen-Image 2.1 Turbo AWQQwen-Image 2.1 Turbo GPTQTimesFM 3.0TimesFM 3.0 AWQTimesFM 3.0 GPTQNemotron 3 DiarizationNemotron 3 Diarization AWQNemotron 3 Diarization GPTQOpenJev GGUF AWQOpenJev GGUF GPTQSmolLM2 360M InstructSmolLM2 360M Instruct AWQSmolLM2 360M Instruct GPTQQwen3-VL 8B InstructQwen3-VL 8B Instruct AWQQwen3-VL 8B Instruct GPTQQwen 3 8B GGUFQwen 3 8B GGUF AWQQwen 3 8B GGUF GPTQOPT 125MOPT 125M AWQOPT 125M GPTQGranite 4.0 H TinyGranite 4.0 H Tiny AWQGranite 4.0 H Tiny GPTQ

Example commands

Serve a model

badgr serve Qwen/Qwen3-32B \
  --gpu A100 40GB \
  --max-cost 10

Starts an OpenAI-compatible endpoint with max-cost protection

Run a custom job

badgr run "python job.py" \
  --gpu A100 40GB \
  --max-cost 10 \
  --max-runtime 120

Runs a script with cost and time limits, logs streamed back

Estimated pricing & availability

Checking last-seen pricing...

Prices below are last-seen estimates. Badgr rechecks live price and availability before launch — final cost depends on selected route, runtime, region, and workload duration.

Loading routes...

Frequently asked questions

Can I rent A100 40GB for non-AI workloads?

Yes. Badgr can run AI workloads, CUDA jobs, scripts, containers, batch jobs, and other GPU workloads. A100 40GB is popular for AI because that's the most common use case, but you can run any CUDA-compatible workload.

Can I run custom Docker containers on A100 40GB?

Yes. Use badgr run with a custom Docker image, or submit a container through the Compute API. Badgr handles orchestration, networking, logs, and teardown.

Does Badgr enforce max-cost limits?

Yes. Set a max-cost budget before launch. If the job would exceed it, Badgr stops the job and issues a receipt. This protects you from runaway costs.

Is A100 40GB availability guaranteed?

No. Badgr rechecks availability immediately before launch. If the route is no longer available, we notify you and suggest alternatives. This is why we show live pricing.

Do I get logs and receipts?

Yes. Stream logs in real-time while the job runs. After completion, Badgr issues a detailed receipt with compute time, cost, and resource usage.

Can I serve models on A100 40GB?

Yes. Use badgr serve to launch an OpenAI-compatible endpoint. Set a max-cost budget and Badgr auto-manages provisioning, health checks, logs, and teardown.

Ready to run on A100 40GB?

Launch from the dashboard, CLI, or Compute API. Max-cost protection included.