# qwen3.8-27b

> The Qwen3.8 27B native vision-language dense model builds upon the 3.6-27B version, with key improvements in coding and office productivity capabilities across both text and visual modalities. It enables more reliable end-to-end completion of complex tasks, delivering consistently trustworthy results.

## Inference Service Provider <span id="h-3827b00010" /> <span id="sec-3827b0000f" />

The inference service provider for `qwen3.8-27b` is Alibaba Cloud Model Studio.

## Model Capabilities <span id="h-3827b00013" /> <span id="sec-3827b00012" />

<table style={{ display: "table", tableLayout: "fixed", width: "100%" }}><colgroup><col style={{ width: "25%" }} /><col style={{ width: "25%" }} /><col style={{ width: "25%" }} /><col style={{ width: "25%" }} /></colgroup><thead><tr><th><p>Capability</p></th><th><p>Support</p></th><th><p>Capability</p></th><th><p>Support</p></th></tr></thead><tbody><tr><td><p>Input Modality</p></td><td><p><strong>Image</strong> <strong>Text</strong> <strong>Video</strong></p></td><td><p>Output Modality</p></td><td><p><strong>Text</strong></p></td></tr><tr><td><p>Model Experience</p></td><td><p>Supported</p></td><td><p>Function Calling</p></td><td><p>Supported</p></td></tr><tr><td><p>Structured Outputs</p></td><td><p>Supported</p></td><td><p>Web Search</p></td><td><p>Supported</p></td></tr><tr><td><p>Prefix Completion</p></td><td><p>Supported</p></td><td><p>Context Caching</p></td><td><p>Supported</p></td></tr><tr><td><p>Batch Inference</p></td><td><p>Unsupported</p></td><td><p>Fine-tuning</p></td><td><p>Unsupported</p></td></tr></tbody></table>

## Context Limits <span id="h-3827b00051" /> <span id="sec-3827b00050" />

<table style={{ display: "table", tableLayout: "fixed", width: "100%" }}><colgroup><col style={{ width: "25%" }} /><col style={{ width: "25%" }} /><col style={{ width: "25%" }} /><col style={{ width: "25%" }} /></colgroup><thead><tr><th><p>Parameter</p></th><th><p>Value</p></th><th><p>Parameter</p></th><th><p>Value</p></th></tr></thead><tbody><tr><td><p>Max Input Length</p></td><td><p>991808</p></td><td><p>Max Output Length</p></td><td><p>131072</p></td></tr><tr><td><p>Max Input Length (Thinking Mode)</p></td><td><p>983616</p></td><td><p>Max Output Length (Thinking Mode)</p></td><td><p>131072</p></td></tr><tr><td><p>Context Window</p></td><td><p>1000000</p></td><td><p>Max Chain-of-Thought Length</p></td><td><p>262144</p></td></tr></tbody></table>

> The supported model length may vary depending on different combinations of API input parameters.

## Pricing <span id="h-3827b000d5" /> <span id="sec-3827b000d4" />

This page only shows the original pricing for model API calls, excluding any limited-time promotions. Visit [Model Studio Console](https://modelstudio.console.alibabacloud.com/ap-southeast-1/model/market) for promotional offers.

<Tabs>
  <Tab title="China (Beijing)">
    <table style={{ display: "table", tableLayout: "fixed", width: "100%" }}><colgroup><col style={{ width: "33.333333%" }} /><col style={{ width: "33.333333%" }} /><col style={{ width: "33.333333%" }} /></colgroup><thead><tr><th><p>Billing Item</p></th><th><p>Price (USD)</p></th><th><p>Unit</p></th></tr></thead><tbody><tr><td><p>Input</p></td><td><p>0.424</p></td><td><p>Per 1M tokens</p></td></tr><tr><td><p>Output</p></td><td><p>1.696</p></td><td><p>Per 1M tokens</p></td></tr><tr><td><p>Input(Implicit Cache)</p></td><td><p>0.085</p></td><td><p>Per 1M tokens</p></td></tr><tr><td><p>Explicit Cache Creation</p></td><td><p>0.53</p></td><td><p>Per 1M tokens</p></td></tr><tr><td><p>Explicit Cache Read</p></td><td><p>0.042</p></td><td><p>Per 1M tokens</p></td></tr></tbody></table>
  </Tab>

  <Tab title="Singapore">
    Scope: International

    <table style={{ display: "table", tableLayout: "fixed", width: "100%" }}><colgroup><col style={{ width: "33.333333%" }} /><col style={{ width: "33.333333%" }} /><col style={{ width: "33.333333%" }} /></colgroup><thead><tr><th><p>Billing Item</p></th><th><p>Price (USD)</p></th><th><p>Unit</p></th></tr></thead><tbody><tr><td><p>Input</p></td><td><p>0.5</p></td><td><p>Per 1M tokens</p></td></tr><tr><td><p>Output</p></td><td><p>3</p></td><td><p>Per 1M tokens</p></td></tr><tr><td><p>Input(Implicit Cache)</p></td><td><p>0.1</p></td><td><p>Per 1M tokens</p></td></tr><tr><td><p>Explicit Cache Creation</p></td><td><p>0.625</p></td><td><p>Per 1M tokens</p></td></tr><tr><td><p>Explicit Cache Read</p></td><td><p>0.05</p></td><td><p>Per 1M tokens</p></td></tr></tbody></table>
  </Tab>
</Tabs>

## Rate Limits <span id="h-3827b0015b" /> <span id="sec-3827b0015a" />

<Tabs>
  <Tab title="China (Beijing)">
    This region uses dynamic rate limiting. The TPM limit is tiered by your monthly Model Studio spend, and the RPM limit is high enough that normal use does not trigger throttling. For the tier values, see [Dynamic rate limiting](/help/en/model-studio/quota-management).
  </Tab>

  <Tab title="Singapore">
    Scope: International

    This region uses dynamic rate limiting. The TPM limit is tiered by your monthly Model Studio spend, and the RPM limit is high enough that normal use does not trigger throttling. For the tier values, see [Dynamic rate limiting](/help/en/model-studio/quota-management).
  </Tab>
</Tabs>