Serve on Badgr

Serve Qwen3-0.6B

Qwen3-0.6B served as an OpenAI-compatible endpoint on Badgr GPU capacity.

GPU recommendation auto-generated from HuggingFace model data — not hand-verified by Badgr.

Recommended GPU

RTX 4090 24GB

Available regions

United States, Europe

Estimated cost range

$0.17 – $0.17/hr

Startup time

2–5 min

Also runs well on L40S 48GB, A100 40GB.

OpenAI-compatible API example

from openai import OpenAI

client = OpenAI(base_url="https://api.aibadgr.com/v1", api_key="YOUR_BADGR_API_KEY")

response = client.chat.completions.create(
    model="qwen3-0-6b",
    messages=[{"role": "user", "content": "Hello!"}],
)
print(response.choices[0].message.content)

Run command

badgr serve qwen3-0-6b \
  --gpu RTX 4090 \
  --max-cost 10

Context length: context length varies by deployment · Model ID: Qwen/Qwen3-0.6B

Recommended Badgr routes

RTX 4090 24GB · United States

Available

from $0.17/hr

Best for

ComfyUI, Batch Inference, Qwen 2.5 7B Instruct

Startup: 2–5 min

Reliability: Standard

Last checked: just now

RTX 4090 24GB · United States

Available

from $0.17/hr

Best for

ComfyUI, Batch Inference, Qwen 2.5 7B Instruct

Startup: 2–5 min

Reliability: Standard

Last checked: just now

RTX 4090 24GB · United States

Available

from $0.17/hr

Best for

ComfyUI, Batch Inference, Qwen 2.5 7B Instruct

Startup: 2–5 min

Reliability: Standard

Last checked: just now

RTX 4090 24GB · Europe

Available

from $0.17/hr

Best for

ComfyUI, Batch Inference, Qwen 2.5 7B Instruct

Startup: 2–5 min

Reliability: Standard

Last checked: just now

RTX 4090 24GB · Europe

Available

from $0.17/hr

Best for

ComfyUI, Batch Inference, Qwen 2.5 7B Instruct

Startup: 2–5 min

Reliability: Standard

Last checked: just now

RTX 4090 24GB · Europe

Available

from $0.17/hr

Best for

ComfyUI, Batch Inference, Qwen 2.5 7B Instruct

Startup: 2–5 min

Reliability: Standard

Last checked: just now