> ## Documentation Index
> Fetch the complete documentation index at: https://veniceai-feat-models-redesign.mintlify.site/llms.txt
> Use this file to discover all available pages before exploring further.

# Qwen 3.8 27B API

> Qwen 3.8 27B API on Venice: 256K context, $0.45 input and $3.20 output per 1M tokens. Private and end-to-end encrypted variants.

export const HubMount = ({view, children, ...props}) => {
  const [hub, setHub] = useState(null);
  const [failed, setFailed] = useState(false);
  useEffect(() => {
    let alive = true;
    const w = window;
    if (!w.__veniceModelHub) {
      const urls = w.location.hostname === 'localhost' ? ['http://localhost:3333/data/model-hub.bundle.json', '/data/model-hub.bundle.json'] : ['/data/model-hub.bundle.json'];
      const load = i => fetch(urls[i], {
        cache: 'no-cache'
      }).then(res => {
        if (!res.ok) throw new Error(`bundle ${res.status}`);
        return res.json();
      }).catch(err => i + 1 < urls.length ? load(i + 1) : Promise.reject(err));
      const Frag = <></>.type;
      const h = (type, props, ...kids) => {
        const T = type;
        const {key, ...rest} = props || ({});
        if (!kids.length) return <T key={key} {...rest} />;
        if (kids.length === 1) return <T key={key} {...rest}>{kids[0]}</T>;
        return <T key={key} {...rest}>{kids.map((kid, i) => <Frag key={i}>{kid}</Frag>)}</T>;
      };
      w.__veniceModelHub = load(0).then(bundle => new Function(`return (${bundle.code})`)()({
        h,
        Fragment: Frag,
        useState,
        useEffect,
        useRef,
        useMemo,
        useCallback
      }));
    }
    w.__veniceModelHub.then(instance => {
      if (alive) setHub(instance);
    }).catch(() => {
      w.__veniceModelHub = null;
      if (alive) setFailed(true);
    });
    return () => {
      alive = false;
    };
  }, []);
  const View = hub ? hub[view] : null;
  if (View) return <View {...props}>{children}</View>;
  if (failed) {
    return <div className="vx-mount is-failed">
        <p className="vx-mount-note">The interactive model catalog could not load. The full data is below.</p>
        {children}
      </div>;
  }
  return <div className="vx-mount" aria-busy="true">
      <div className="vx-mount-skeleton" aria-hidden="true"><span /><span /><span /></div>
      <div className="vx-mount-source">{children}</div>
    </div>;
};

<HubMount view="ModelPage" data={{"family":{"slug":"qwen-3-8-27b","name":"Qwen 3.8 27B","modality":"text","task":"chat","provider":"alibaba","description":"Qwen 3.8 27B is a native vision-language dense model with 27B parameters. It improves coding, professional work, research, and long-horizon agentic tasks, with flexible thinking control and image and video understanding. It supports a native 262K-token context window.","primary":"qwen-3-8-27b","variants":["qwen-3-8-27b","e2ee-qwen3-8-27b"],"created":1786924800,"updated":1788652800,"privacy":["private","e2ee"],"uncensored":true,"openWeights":true,"traits":["default_vision"]},"models":[{"id":"qwen-3-8-27b","name":"Qwen 3.8 27B","type":"text","modality":"text","task":"chat","variant":"standard","provider":"alibaba","created":1786924800,"description":"Qwen 3.8 27B is a native vision-language dense model with 27B parameters. It improves coding, professional work, research, and long-horizon agentic tasks, with flexible thinking control and image and video understanding. It supports a native 262K-token context window.","source":"https://huggingface.co/Qwen/Qwen3.8-27B-FP8","privacy":"private","uncensored":true,"beta":true,"traits":["default_vision"],"openWeights":true,"text":{"context":262144,"maxOutput":65536,"quantization":"fp8","reasoning":{"supported":true,"effort":["none","low","medium","xhigh"],"defaultEffort":"xhigh"},"caps":{"tools":true,"structured":true,"vision":true,"maxImages":10,"videoInput":true,"webSearch":true,"code":true}},"pricing":{"input":0.45,"output":3.2,"blended":1.1375},"headline":{"value":1.1375,"unit":"per 1M tokens","basis":"blended"},"endpoints":[{"id":"chat","method":"POST","path":"/chat/completions","name":"Chat Completions","status":"stable","recommended":true},{"id":"responses","method":"POST","path":"/responses","name":"Responses","status":"alpha"}],"family":"qwen-3-8-27b"},{"id":"e2ee-qwen3-8-27b","name":"Qwen 3.8 27B","type":"text","modality":"text","task":"chat","variant":"e2ee","provider":"alibaba","created":1788652800,"description":"Qwen 3.8 27B running in a Trusted Execution Environment (TEE). A multimodal model with 262K context supporting text and image input. Hardware attestation evidence is available for independent verification of enclave identity and configuration.","source":"https://huggingface.co/Qwen/Qwen3.8-27B","privacy":"e2ee","openWeights":true,"text":{"context":262144,"maxOutput":8192,"reasoning":{"supported":true},"caps":{"tools":true,"vision":true,"maxImages":1,"webSearch":true,"code":true,"e2ee":true,"tee":true}},"pricing":{"input":0.47,"output":3.53,"cacheRead":0.047,"blended":1.235},"headline":{"value":1.235,"unit":"per 1M tokens","basis":"blended"},"endpoints":[{"id":"chat","method":"POST","path":"/chat/completions","name":"Chat Completions","status":"stable","recommended":true},{"id":"responses","method":"POST","path":"/responses","name":"Responses","status":"alpha","note":"E2EE models require /chat/completions with E2EE headers."}],"family":"qwen-3-8-27b"}],"related":{"similar":[{"slug":"aion-3-5-mini","name":"Aion 3.5 Mini","provider":"aion","modality":"text","privacy":["anonymized"],"created":1790121600,"headline":{"value":1.09375,"unit":"per 1M tokens","basis":"blended"},"variants":1},{"slug":"aion-3-0-mini","name":"Aion 3.0 Mini","provider":"aion","modality":"text","privacy":["anonymized"],"created":1783468800,"headline":{"value":1.09375,"unit":"per 1M tokens","basis":"blended"},"variants":1},{"slug":"gemini-3-5-flash-lite","name":"Gemini 3.5 Flash-Lite","provider":"google","modality":"text","privacy":["anonymized"],"created":1783555200,"headline":{"value":1.0625,"unit":"per 1M tokens","basis":"blended"},"variants":1},{"slug":"seed-2-1-turbo","name":"Seed 2.1 Turbo","provider":"bytedance","modality":"text","privacy":["anonymized"],"created":1782604800,"headline":{"value":1.25,"unit":"per 1M tokens","basis":"blended"},"variants":1}],"versions":[{"slug":"qwen-3-6-27b","name":"Qwen 3.6 27B","provider":"alibaba","modality":"text","privacy":["private"],"created":1776988800,"headline":{"value":1.05625,"unit":"per 1M tokens","basis":"blended"},"variants":1},{"slug":"qwen-2-5-7b","name":"Qwen 2.5 7B","provider":"alibaba","modality":"text","privacy":["e2ee"],"created":1773792000,"headline":{"value":0.07,"unit":"per 1M tokens","basis":"blended"},"variants":1},{"slug":"qwen-3-5-9b","name":"Qwen 3.5 9B","provider":"alibaba","modality":"text","privacy":["private"],"created":1772668800,"headline":{"value":0.1125,"unit":"per 1M tokens","basis":"blended"},"variants":1},{"slug":"qwen-3-5-397b","name":"Qwen 3.5 397B","provider":"alibaba","modality":"text","privacy":["anonymized"],"created":1771200000,"headline":{"value":1.6875,"unit":"per 1M tokens","basis":"blended"},"variants":1}]},"providers":{"alibaba":{"slug":"alibaba","name":"Alibaba Qwen","logo":"/images/icons/models/qwen.svg"},"aion":{"slug":"aion","name":"Aion Labs","logo":"/images/icons/models/aionlabs.svg"},"google":{"slug":"google","name":"Google","logo":"/images/icons/models/google.svg"},"bytedance":{"slug":"bytedance","name":"ByteDance Seed","logo":"/images/icons/models/bytedance.svg"}},"faq":[{"q":"How much does the Qwen 3.8 27B API cost?","a":"$0.45 per 1M input tokens and $3.20 per 1M output tokens. The E2EE variant costs $0.47 input and $3.53 output. Prices are in USD and can be paid in DIEM at parity."},{"q":"What is the Qwen 3.8 27B model ID?","a":"Use `qwen-3-8-27b` as the `model` parameter. Other variants: `e2ee-qwen3-8-27b` (E2EE)."},{"q":"Is the Qwen 3.8 27B API private?","a":"The Standard variant is private: requests run on infrastructure Venice controls with zero data retention, and prompts and outputs are never stored or used for training. The E2EE variant is end-to-end encrypted: prompts are encrypted on your device and decrypted only inside an attested hardware enclave, so neither Venice nor the GPU provider can read them."},{"q":"What is the context window of Qwen 3.8 27B?","a":"256K tokens of context, with up to 64K output tokens per response."},{"q":"What does Qwen 3.8 27B support?","a":"Qwen 3.8 27B supports function calling, structured outputs, reasoning, image input and web search. Reasoning effort is adjustable with `reasoning_effort`: none, low, medium and xhigh (default xhigh)."},{"q":"Which endpoint does the Qwen 3.8 27B API use?","a":"Call `POST /chat/completions`. `/responses` (Alpha) is also supported."}]}} />

<div className="vx-static">
  <Accordion title="Plain-text specification">
    # Qwen 3.8 27B API

    Qwen 3.8 27B is a large language model by Alibaba Qwen, available on the Venice API as `qwen-3-8-27b`. Private and end-to-end encrypted variants are available.

    Qwen 3.8 27B is a native vision-language dense model with 27B parameters. It improves coding, professional work, research, and long-horizon agentic tasks, with flexible thinking control and image and video understanding. It supports a native 262K-token context window.

    ## Qwen 3.8 27B API pricing

    | Model ID | Variant | Privacy | Price |
    | - | - | - | - |
    | `qwen-3-8-27b` | Standard | Private | $0.45 input / $3.20 output per 1M tokens |
    | `e2ee-qwen3-8-27b` | E2EE | End-to-end encrypted | $0.47 input / $3.53 output per 1M tokens |

    ## Qwen 3.8 27B specifications

    | Spec | Value |
    | - | - |
    | Provider | Alibaba Qwen |
    | Released | Aug 17, 2026 |
    | Privacy | Private, End-to-end encrypted |
    | Open weights | Yes |
    | Context window | 256K tokens |
    | Max output | 64K tokens |
    | Reasoning effort | none, low, medium, xhigh |
    | Served precision | FP8 |

    ## How to use the Qwen 3.8 27B API

    Send requests to `POST https://api.venice.ai/api/v1/chat/completions` with `"model": "qwen-3-8-27b"` and your API key.

    ```bash theme={null}
    curl https://api.venice.ai/api/v1/chat/completions \
      -H "Authorization: Bearer $VENICE_API_KEY" \
      -H "Content-Type: application/json" \
      -d '{
        "model": "qwen-3-8-27b",
        "messages": [{ "role": "user", "content": "Explain TEE attestation in two sentences." }],
        "reasoning_effort": "xhigh"
      }'
    ```

    ## Qwen 3.8 27B API FAQ

    ### How much does the Qwen 3.8 27B API cost?

    $0.45 per 1M input tokens and $3.20 per 1M output tokens. The E2EE variant costs $0.47 input and $3.53 output. Prices are in USD and can be paid in DIEM at parity.

    ### What is the Qwen 3.8 27B model ID?

    Use `qwen-3-8-27b` as the `model` parameter. Other variants: `e2ee-qwen3-8-27b` (E2EE).

    ### Is the Qwen 3.8 27B API private?

    The Standard variant is private: requests run on infrastructure Venice controls with zero data retention, and prompts and outputs are never stored or used for training. The E2EE variant is end-to-end encrypted: prompts are encrypted on your device and decrypted only inside an attested hardware enclave, so neither Venice nor the GPU provider can read them.

    ### What is the context window of Qwen 3.8 27B?

    256K tokens of context, with up to 64K output tokens per response.

    ### What does Qwen 3.8 27B support?

    Qwen 3.8 27B supports function calling, structured outputs, reasoning, image input and web search. Reasoning effort is adjustable with `reasoning_effort`: none, low, medium and xhigh (default xhigh).

    ### Which endpoint does the Qwen 3.8 27B API use?

    Call `POST /chat/completions`. `/responses` (Alpha) is also supported.

    ## Related models

    * [Qwen 3.6 27B API](/models/qwen-3-6-27b): $0.33 input / $3.25 output per 1M tokens
    * [Qwen 2.5 7B API](/models/qwen-2-5-7b): $0.05 input / $0.13 output per 1M tokens
    * [Qwen 3.5 9B API](/models/qwen-3-5-9b): $0.10 input / $0.15 output per 1M tokens
    * [Qwen 3.5 397B API](/models/qwen-3-5-397b): $0.75 input / $4.50 output per 1M tokens
    * [Aion 3.5 Mini API](/models/aion-3-5-mini): $0.88 input / $1.75 output per 1M tokens
    * [Aion 3.0 Mini API](/models/aion-3-0-mini): $0.88 input / $1.75 output per 1M tokens
    * [Gemini 3.5 Flash-Lite API](/models/gemini-3-5-flash-lite): $0.38 input / $3.13 output per 1M tokens
    * [Seed 2.1 Turbo API](/models/seed-2-1-turbo): $0.63 input / $3.13 output per 1M tokens
  </Accordion>
</div>


This documentation is built and hosted on [Mintlify](https://mintlify.com), a developer documentation platform.