Qwen: Qwen3.5 35B A3B
Chatbailian/qwen3.5-35b-a3bQwen3.5-35B-A3B is a hybrid-architecture vision-language model from Alibaba's Qwen3.5 series, combining linear attention with sparse MoE. With only 3B active parameters, it delivers inference efficiency close to Qwen3.5-122B-A10B. Context: 256K tokens, output: 64K. Accessible via the OpenAI-compatible protocol.
Context Window
256K
Max Output Tokens
64K
Released
2026-02-23
Capabilities
VisionFunction CallingReasoningPrompt CachingWeb SearchVideo Input
Available Providers
Aliyun
Supported Protocols
openai
Providers
Aliyun
Input Tokens
$0.29/M
Output Tokens
$1.83/M
Cache Read
$0.29/M
Web Search
$0.01/R
Protocols
openai
/v1/chat/completionsCode Examples
from openai import OpenAIclient = OpenAI(base_url="https://api.ofox.io/v1",api_key="YOUR_OFOX_API_KEY",)response = client.chat.completions.create(model="bailian/qwen3.5-35b-a3b",messages=[{"role": "user", "content": "Hello!"}],)print(response.choices[0].message.content)
Benchmarks
LMArena ↗Evaluated as qwen3.5-35b-a3b
Qwen: Qwen3.5 35B A3B scores 1396 in the Overall category of the LMArena text leaderboard (style control), ranking #132 of 374 models based on 29,192 human preference votes (updated 2026-07-12).
| Category | Arena Score | 95% CI | Rank | Votes |
|---|---|---|---|---|
| Overall | 1396 | 1391–1400 | #132 of 374 | 29,192 |
| Hard Prompts | 1413 | 1408–1418 | #138 of 374 | 18,360 |
| Coding | 1435 | 1427–1442 | #139 of 369 | 7,985 |
| Math | 1399 | 1385–1413 | #125 of 362 | 1,758 |
| Creative Writing | 1343 | 1334–1353 | #152 of 372 | 4,483 |
| Instruction Following | 1388 | 1381–1395 | #128 of 374 | 9,312 |
| Chinese | 1461 | 1445–1476 | #93 of 344 | 1,608 |
Source: LMArena · CC BY 4.0 · Updated 2026-07-12 · Methodology ↗ · Ranks compare models within each category of the LMArena text leaderboard (style control). Scores come from third-party human preference evaluations, not from OFOX.
Related Models
Frequently Asked Questions
Qwen: Qwen3.5 35B A3B on Ofox.ai costs $0.29/M per million input tokens and $1.83/M per million output tokens. Pay-as-you-go, no monthly fees.
Qwen: Qwen3.5 35B A3B supports a context window of 256K tokens with max output of 64K tokens, allowing you to process large documents and maintain long conversations.
Simply set your base URL to https://api.ofox.io/v1 and use your Ofox API key. The API is OpenAI-compatible — just change the base URL and API key in your existing code.
Qwen: Qwen3.5 35B A3B supports the following capabilities: Vision, Function Calling, Reasoning, Prompt Caching, Web Search, Video Input. Access all features through the Ofox.ai unified API.