20% off all GPT ๐ŸŽ‰ 30% off DeepSeekLearn more

Qwen: Qwen3.8 27B

Chat
qwen/qwen3.8-27b

Qwen3.8 27B is the natively vision-language dense model in Alibaba Cloud Bailian's Qwen3.8 series, released on 2026-08-19. It handles image and video understanding and supports deep reasoning (reasoning_content) that can be toggled per request, and focuses its gains over Qwen3.6 27B on coding and office scenarios in both text and visual modalities. It also supports tool use, prompt caching, and web search. Context window: 1.13M tokens, output: 131K. As a dense model rather than a Mixture-of-Experts one, it offers more predictable latency for interactive multimodal workloads. Available via OpenAI and Anthropic protocols through Ofox.

Context Window
1M
Max Output Tokens
131K
Released
2026-08-19
Capabilities
VisionFunction CallingReasoningPrompt CachingWeb SearchVideo Input
Available Providers
AliyunAlibaba Cloud
Supported Protocols
openaianthropic

Providers

Alibaba Cloudprovider.type: "alicloud"
Input Tokens
$0.5/M
Output Tokens
$1.71/M
Cache Read
$0.043/M
Cache Write
$0.63/M
Web Search
$0.01/R
Protocols
openai/v1/chat/completions/v1/responses
anthropic
Aliyunprovider.type: "aliyun"
Input Tokens
$0.5/M
Output Tokens
$1.71/M
Cache Read
$0.043/M
Cache Write
$0.63/M
Web Search
$0.01/R
Protocols
openai/v1/chat/completions/v1/responses
anthropic

Code Examples

from openai import OpenAI
client = OpenAI(
base_url="https://api.ofox.run/v1",
api_key="YOUR_OFOX_API_KEY",
)
response = client.chat.completions.create(
model="qwen/qwen3.8-27b",
messages=[
{"role": "user", "content": "Hello!"}
],
)
print(response.choices[0].message.content)

Uptime & Status

Frequently Asked Questions

Qwen: Qwen3.8 27B on Ofox.ai costs $0.5/M per million input tokens and $1.71/M per million output tokens. Pay-as-you-go, no monthly fees.