Serve on Badgr

Serve Sao10K: Llama 3 8B Lunaris

Sao10K: Llama 3 8B Lunaris served as an OpenAI-compatible endpoint on Badgr GPU capacity.

GPU recommendation auto-generated from OpenRouter model data — not hand-verified by Badgr.

Recommended GPU

RTX 4090 24GB

Available regions

United States, Europe

Estimated cost range

$0.17 – $0.17/hr

Startup time

2–5 min

Also runs well on L40S 48GB, A100 40GB.

OpenAI-compatible API example

from openai import OpenAI

client = OpenAI(base_url="https://api.aibadgr.com/v1", api_key="YOUR_BADGR_API_KEY")

response = client.chat.completions.create(
    model="l3-lunaris-8b",
    messages=[{"role": "user", "content": "Hello!"}],
)
print(response.choices[0].message.content)

Run command

badgr serve l3-lunaris-8b \
  --gpu RTX 4090 \
  --max-cost 10

Context length: 8K tokens · Model ID: sao10k/l3-lunaris-8b

Recommended Badgr routes

RTX 4090 24GB · United States

Available

from $0.17/hr

Best for

ComfyUI, Batch Inference, Qwen 2.5 7B Instruct

Startup: 2–5 min

Reliability: Standard

Last checked: just now

RTX 4090 24GB · United States

Available

from $0.17/hr

Best for

ComfyUI, Batch Inference, Qwen 2.5 7B Instruct

Startup: 2–5 min

Reliability: Standard

Last checked: just now

RTX 4090 24GB · United States

Available

from $0.17/hr

Best for

ComfyUI, Batch Inference, Qwen 2.5 7B Instruct

Startup: 2–5 min

Reliability: Standard

Last checked: just now

RTX 4090 24GB · Europe

Available

from $0.17/hr

Best for

ComfyUI, Batch Inference, Qwen 2.5 7B Instruct

Startup: 2–5 min

Reliability: Standard

Last checked: just now

RTX 4090 24GB · Europe

Available

from $0.17/hr

Best for

ComfyUI, Batch Inference, Qwen 2.5 7B Instruct

Startup: 2–5 min

Reliability: Standard

Last checked: just now

RTX 4090 24GB · Europe

Available

from $0.17/hr

Best for

ComfyUI, Batch Inference, Qwen 2.5 7B Instruct

Startup: 2–5 min

Reliability: Standard

Last checked: just now