← All models
OnlineItera Hosted

Qwen3.8-27B

Qwen · OpenAI-compatible API

qwen/qwen3.8-27b

Free

Input$0
Cached Input$0
Output$0
Context
32K Free
Tool Calling
✓
Reasoning
✓
Vision
✓
Streaming
✓
OpenAI-compatible API
import os
from openai import OpenAI

client = OpenAI(
    base_url="https://api.iteracompute.com/v1",
    api_key=os.environ["ITERACOMPUTE_API_KEY"],
)
response = client.chat.completions.create(
    model="qwen/qwen3.8-27b",
    messages=[{"role": "user", "content": "Hello!"}],
)
print(response.choices[0].message.content)

Overview

Qwen3.8-27B is the free entry model on IteraCompute GPUs, with 32K free context, tool calling and streaming. OpenAI-compatible API access is now available for free.

Pricing

Free · Itera Hosted. Input, cached input and output are $0 per 1M tokens. Fair-use rate limits apply; no SLA on free capacity.

Input$0
Cached Input$0
Output$0

API

Base URL
https://api.iteracompute.com/v1
Endpoint
POST /v1/chat/completions
Model ID
qwen/qwen3.8-27b
Authentication
Authorization: Bearer ITERACOMPUTE_API_KEY

View Docs

Use Cases

Prototype chat experiences, learn tool calling and test streaming integrations with free capacity.

FAQ

How do I use Qwen3.8-27B with the OpenAI SDK?

Set base_url to https://api.iteracompute.com/v1 and model to qwen/qwen3.8-27b. Authenticate with your IteraCompute API key.

What does Qwen3.8-27B cost?

Input $0, cached input $0, output $0 per 1M tokens, in USD. Free capacity is subject to fair-use rate limits and has no SLA.

What context length is available?

32,768 tokens. This is the free context limit.