> ## Documentation Index
> Fetch the complete documentation index at: https://veniceai-feat-models-redesign.mintlify.site/llms.txt
> Use this file to discover all available pages before exploring further.

# Qwen 3.6 27B API

> Qwen 3.6 27B API on Venice: 250K context, $0.33 input and $3.25 output per 1M tokens. Private, with zero data retention. Pricing, specs and code examples.

export const HubMount = ({view, children, ...props}) => {
  const [hub, setHub] = useState(null);
  const [failed, setFailed] = useState(false);
  useEffect(() => {
    let alive = true;
    const w = window;
    if (!w.__veniceModelHub) {
      const urls = w.location.hostname === 'localhost' ? ['http://localhost:3333/data/model-hub.bundle.json', '/data/model-hub.bundle.json'] : ['/data/model-hub.bundle.json'];
      const load = i => fetch(urls[i], {
        cache: 'no-cache'
      }).then(res => {
        if (!res.ok) throw new Error(`bundle ${res.status}`);
        return res.json();
      }).catch(err => i + 1 < urls.length ? load(i + 1) : Promise.reject(err));
      const Frag = <></>.type;
      const h = (type, props, ...kids) => {
        const T = type;
        const {key, ...rest} = props || ({});
        if (!kids.length) return <T key={key} {...rest} />;
        if (kids.length === 1) return <T key={key} {...rest}>{kids[0]}</T>;
        return <T key={key} {...rest}>{kids.map((kid, i) => <Frag key={i}>{kid}</Frag>)}</T>;
      };
      w.__veniceModelHub = load(0).then(bundle => new Function(`return (${bundle.code})`)()({
        h,
        Fragment: Frag,
        useState,
        useEffect,
        useRef,
        useMemo,
        useCallback
      }));
    }
    w.__veniceModelHub.then(instance => {
      if (alive) setHub(instance);
    }).catch(() => {
      w.__veniceModelHub = null;
      if (alive) setFailed(true);
    });
    return () => {
      alive = false;
    };
  }, []);
  const View = hub ? hub[view] : null;
  if (View) return <View {...props}>{children}</View>;
  if (failed) {
    return <div className="vx-mount is-failed">
        <p className="vx-mount-note">The interactive model catalog could not load. The full data is below.</p>
        {children}
      </div>;
  }
  return <div className="vx-mount" aria-busy="true">
      <div className="vx-mount-skeleton" aria-hidden="true"><span /><span /><span /></div>
      <div className="vx-mount-source">{children}</div>
    </div>;
};

<HubMount view="ModelPage" data={{"family":{"slug":"qwen-3-6-27b","name":"Qwen 3.6 27B","modality":"text","task":"chat","provider":"alibaba","description":"The Qwen 3.6 27B native vision-language dense model builds upon the 3.5-27B version, with key improvements in agentic coding capabilities and enhanced STEM reasoning and inference skills. In the vision modality, it demonstrates significant advances in spatial intelligence, object localization, and detection, while video understanding, document OCR, and visual agent capabilities continue to improve steadily.","primary":"qwen3-6-27b","variants":["qwen3-6-27b"],"created":1776988800,"updated":1776988800,"privacy":["private"],"openWeights":true},"models":[{"id":"qwen3-6-27b","name":"Qwen 3.6 27B","type":"text","modality":"text","task":"chat","variant":"standard","provider":"alibaba","created":1776988800,"description":"The Qwen 3.6 27B native vision-language dense model builds upon the 3.5-27B version, with key improvements in agentic coding capabilities and enhanced STEM reasoning and inference skills. In the vision modality, it demonstrates significant advances in spatial intelligence, object localization, and detection, while video understanding, document OCR, and visual agent capabilities continue to improve steadily.","source":"https://huggingface.co/Qwen/Qwen3.6-27B-FP8","privacy":"private","openWeights":true,"text":{"context":256000,"maxOutput":65536,"quantization":"fp8","reasoning":{"supported":true,"effort":["none","low","medium","high"],"defaultEffort":"low"},"caps":{"tools":true,"structured":true,"vision":true,"maxImages":10,"videoInput":true,"webSearch":true,"code":true},"sampling":{"temperature":1,"top_p":0.95}},"pricing":{"input":0.325,"output":3.25,"blended":1.05625},"headline":{"value":1.05625,"unit":"per 1M tokens","basis":"blended"},"endpoints":[{"id":"chat","method":"POST","path":"/chat/completions","name":"Chat Completions","status":"stable","recommended":true},{"id":"responses","method":"POST","path":"/responses","name":"Responses","status":"alpha"}],"family":"qwen-3-6-27b"}],"related":{"similar":[{"slug":"gemini-3-5-flash-lite","name":"Gemini 3.5 Flash-Lite","provider":"google","modality":"text","privacy":["anonymized"],"created":1783555200,"headline":{"value":1.0625,"unit":"per 1M tokens","basis":"blended"},"variants":1},{"slug":"aion-3-0-mini","name":"Aion 3.0 Mini","provider":"aion","modality":"text","privacy":["anonymized"],"created":1783468800,"headline":{"value":1.09375,"unit":"per 1M tokens","basis":"blended"},"variants":1},{"slug":"glm-4-7","name":"GLM 4.7","provider":"zai","modality":"text","privacy":["private"],"created":1766534400,"headline":{"value":1.075,"unit":"per 1M tokens","basis":"blended"},"variants":1},{"slug":"grok-build-0-1","name":"Grok Build 0.1","provider":"xai","modality":"text","privacy":["private"],"created":1779321600,"headline":{"value":1.25,"unit":"per 1M tokens","basis":"blended"},"variants":1}],"versions":[{"slug":"qwen-3-8-27b","name":"Qwen 3.8 27B","provider":"alibaba","modality":"text","privacy":["private","e2ee"],"created":1786924800,"headline":{"value":1.1375,"unit":"per 1M tokens","basis":"blended"},"variants":2},{"slug":"qwen-2-5-7b","name":"Qwen 2.5 7B","provider":"alibaba","modality":"text","privacy":["e2ee"],"created":1773792000,"headline":{"value":0.07,"unit":"per 1M tokens","basis":"blended"},"variants":1},{"slug":"qwen-3-5-9b","name":"Qwen 3.5 9B","provider":"alibaba","modality":"text","privacy":["private"],"created":1772668800,"headline":{"value":0.1125,"unit":"per 1M tokens","basis":"blended"},"variants":1},{"slug":"qwen-3-5-397b","name":"Qwen 3.5 397B","provider":"alibaba","modality":"text","privacy":["anonymized"],"created":1771200000,"headline":{"value":1.6875,"unit":"per 1M tokens","basis":"blended"},"variants":1}]},"providers":{"alibaba":{"slug":"alibaba","name":"Alibaba Qwen","logo":"/images/icons/models/qwen.svg"},"google":{"slug":"google","name":"Google","logo":"/images/icons/models/google.svg"},"aion":{"slug":"aion","name":"Aion Labs","logo":"/images/icons/models/aionlabs.svg"},"zai":{"slug":"zai","name":"Z.ai","logo":"/images/icons/models/Zhipu.svg"},"xai":{"slug":"xai","name":"xAI","logo":"/images/icons/models/grok.svg"}},"faq":[{"q":"How much does the Qwen 3.6 27B API cost?","a":"$0.33 per 1M input tokens and $3.25 per 1M output tokens. Prices are in USD and can be paid in DIEM at parity."},{"q":"What is the Qwen 3.6 27B model ID?","a":"Use `qwen3-6-27b` as the `model` parameter."},{"q":"Is the Qwen 3.6 27B API private?","a":"Qwen 3.6 27B is private: requests run on infrastructure Venice controls with zero data retention, and prompts and outputs are never stored or used for training."},{"q":"What is the context window of Qwen 3.6 27B?","a":"250K tokens of context, with up to 64K output tokens per response."},{"q":"What does Qwen 3.6 27B support?","a":"Qwen 3.6 27B supports function calling, structured outputs, reasoning, image input and web search. Reasoning effort is adjustable with `reasoning_effort`: none, low, medium and high (default low)."},{"q":"Which endpoint does the Qwen 3.6 27B API use?","a":"Call `POST /chat/completions`. `/responses` (Alpha) is also supported."}]}} />

<div className="vx-static">
  <Accordion title="Plain-text specification">
    # Qwen 3.6 27B API

    Qwen 3.6 27B is a large language model by Alibaba Qwen, available on the Venice API as `qwen3-6-27b`. It runs privately, with zero data retention.

    The Qwen 3.6 27B native vision-language dense model builds upon the 3.5-27B version, with key improvements in agentic coding capabilities and enhanced STEM reasoning and inference skills. In the vision modality, it demonstrates significant advances in spatial intelligence, object localization, and detection, while video understanding, document OCR, and visual agent capabilities continue to improve steadily.

    ## Qwen 3.6 27B API pricing

    | Model ID | Variant | Privacy | Price |
    | - | - | - | - |
    | `qwen3-6-27b` | Standard | Private | $0.33 input / $3.25 output per 1M tokens |

    ## Qwen 3.6 27B specifications

    | Spec | Value |
    | - | - |
    | Provider | Alibaba Qwen |
    | Released | Apr 24, 2026 |
    | Privacy | Private |
    | Open weights | Yes |
    | Context window | 250K tokens |
    | Max output | 64K tokens |
    | Reasoning effort | none, low, medium, high |
    | Served precision | FP8 |

    ## How to use the Qwen 3.6 27B API

    Send requests to `POST https://api.venice.ai/api/v1/chat/completions` with `"model": "qwen3-6-27b"` and your API key.

    ```bash theme={null}
    curl https://api.venice.ai/api/v1/chat/completions \
      -H "Authorization: Bearer $VENICE_API_KEY" \
      -H "Content-Type: application/json" \
      -d '{
        "model": "qwen3-6-27b",
        "messages": [{ "role": "user", "content": "Explain TEE attestation in two sentences." }],
        "reasoning_effort": "low"
      }'
    ```

    ## Qwen 3.6 27B API FAQ

    ### How much does the Qwen 3.6 27B API cost?

    $0.33 per 1M input tokens and $3.25 per 1M output tokens. Prices are in USD and can be paid in DIEM at parity.

    ### What is the Qwen 3.6 27B model ID?

    Use `qwen3-6-27b` as the `model` parameter.

    ### Is the Qwen 3.6 27B API private?

    Qwen 3.6 27B is private: requests run on infrastructure Venice controls with zero data retention, and prompts and outputs are never stored or used for training.

    ### What is the context window of Qwen 3.6 27B?

    250K tokens of context, with up to 64K output tokens per response.

    ### What does Qwen 3.6 27B support?

    Qwen 3.6 27B supports function calling, structured outputs, reasoning, image input and web search. Reasoning effort is adjustable with `reasoning_effort`: none, low, medium and high (default low).

    ### Which endpoint does the Qwen 3.6 27B API use?

    Call `POST /chat/completions`. `/responses` (Alpha) is also supported.

    ## Related models

    * [Qwen 3.8 27B API](/models/qwen-3-8-27b): $0.45 input / $3.20 output per 1M tokens
    * [Qwen 2.5 7B API](/models/qwen-2-5-7b): $0.05 input / $0.13 output per 1M tokens
    * [Qwen 3.5 9B API](/models/qwen-3-5-9b): $0.10 input / $0.15 output per 1M tokens
    * [Qwen 3.5 397B API](/models/qwen-3-5-397b): $0.75 input / $4.50 output per 1M tokens
    * [Gemini 3.5 Flash-Lite API](/models/gemini-3-5-flash-lite): $0.38 input / $3.13 output per 1M tokens
    * [Aion 3.0 Mini API](/models/aion-3-0-mini): $0.88 input / $1.75 output per 1M tokens
    * [GLM 4.7 API](/models/glm-4-7): $0.55 input / $2.65 output per 1M tokens
    * [Grok Build 0.1 API](/models/grok-build-0-1): $1.00 input / $2.00 output per 1M tokens
  </Accordion>
</div>


This documentation is built and hosted on [Mintlify](https://mintlify.com), a developer documentation platform.