Qwen 3.6 35B A3B

CodeVisionReasoningWeb searchFunction calling
Advertised as qwen/qwen3.6-35b-a3bqwen3-6-35b-a3bqwen3.6-35b-a3bQwen3.6-35B-A3B
Quick reference
Qwen 3.6 35B A3B — TLDR
  • - 🧠 Sparse mixture-of-experts: 35B total, ~3B active per token.
  • - 📏 Native 256K-token context window.
  • - 🎯 Tuned for agentic coding, STEM reasoning, and tool use.
  • - 🔧 Built-in function calling and web-search capabilities.
  • - 🔒 Apache-2.0 licensed, self-hostable open weights.
  • - ⚡ Ships as an FP8 checkpoint for efficient inference.
  • - 🆕 Flagship open-weight release of the Qwen3.6 generation.
  • - 🏢 Developed by Alibaba's Qwen team.
💰 Best price on AntSeed
$0.085 / $0.770−43%
per 1M · cheapest in / out
📏 Context
256K tokens
🐜 Sellers
6
advertising on AntSeed
Provider

Alibaba Group is a Chinese multinational technology company founded in 1999 and headquartered in Hangzhou, Zhejiang. Originally built around e-commerce and cloud computing, Alibaba has become one of the most prolific contributors to open-weight AI research, developing the Qwen…

Explore 34 more models by Alibaba Group →
About this model

Qwen 3.6 35B A3B is a text model from Alibaba's Qwen team and the flagship open-weight release of the Qwen3.6 generation. It uses a sparse mixture-of-experts design with 35 billion total parameters but only about 3 billion active per token, and natively supports a 256K-token context window. The weights ship under Apache-2.0, including an official FP8 checkpoint, and the model is tuned for agentic coding, STEM reasoning, and tool use.

Relative to its direct predecessor Qwen 3.5 35B A3B, the model keeps the same 35B-total / 3B-active mixture-of-experts layout, positioning the 3.6 release as a successor within the same size class rather than a parameter scale-up. It also sits alongside the smaller Qwen 3.6 27B within the same family.

The model targets developers building coding agents and tool-using workflows, with function-calling and web-search capabilities and open weights that can be self-hosted. As a compact active-parameter MoE, it aims to combine the throughput of a small model with the capacity of a larger expert pool, making it suitable for latency-sensitive agentic and STEM applications.

View source on GitHub ↗View model card on HuggingFace ↗
Sources
huggingface.coQwen/Qwen3.6-35B-A3B · Hugging Face· huggingface.co

This About section is AI-generated from public sources via VeniceStats + Venice inference, with no human editing. It may contain inaccuracies.

Usage on AntSeed
Tokens served
48.77M
input + output
Requests
1,295
settled calls
Buyers
7
distinct, on this model
Sellers used
4
of 6 advertising
Settled
$9.97
gross USDC, this model
Sellers serving Qwen 3.6 35B A3B (6)compare on the network explorer →
SellerReputation↓RoutingInput $/MCached $/MOutput $/MCategoriesAPI

"Best price" and the seller table are live AntSeed catalog data (advertised $/1M tokens — or $ per generated image for unit-billed image models — not settled amounts). Reputation = buyer trust score (0-100, the AntSeed SDK's own formula). "Routing" = the SDK's default buyer routing (what the VPR desktop app ships with): a trust ≥ 60 gate on the effective reputation, then cheapest-first among routable sellers; live failover state (per-peer cooldowns) is buyer-side runtime and not included. Model knowledge (TLDR, provider, About) via the VeniceStats enrichment layer. Advertised catalog, not the model used in any specific purchase. "Usage on AntSeed" counts only settlements whose buyers share the per-model split on-chain (metadata v2/v3, opt-in), so every usage figure is a lower bound.