Qwen
A RocketRide LLM node that connects Alibaba Cloud Qwen models to a pipeline via the DashScope API.
What it does
Provides Qwen chat completions to the pipeline. Used primarily as an llm invoke connection by agents and other nodes that need an LLM, and can also be used directly via lanes.
Uses LangChain's ChatOpenAI client pointed at DashScope's OpenAI-compatible endpoint. The endpoint is resolved at startup from the base_url field if set, otherwise from the region field. Temperature is fixed at 0, and max_tokens is taken from the profile's modelOutputTokens.
When the node configuration is validated, the node performs a live 1-token test request against the API to verify the key, model, and region actually work. Failures surface as configuration warnings with the provider's error message.
Lanes
| Lane in | Lane out | Description |
|---|---|---|
questions | answers | Send a question directly, receive a generated answer |
Profiles
Default: Qwen Flash (latest) (qwen-flash).
The visible (latest) stable aliases always resolve to DashScope's current snapshot for their tier, so they do not go stale as new generations ship — prefer these unless you need to pin a specific release, which the first four collapsed profiles do.
| Profile | Model | Context | Output |
|---|---|---|---|
qwen-flash (default) | qwen-flash | 131,072 | 4,096 |
qwen-plus | qwen-plus | 1,000,000 | 32,768 |
qwen-max | qwen-max | 32,768 | 8,192 |
qwen-turbo | qwen-turbo | 131,072 | 8,192 |
View 45 more models
| Profile | Model | Context | Output |
|---|---|---|---|
qwen3-7-max | qwen3.7-max | 1,000,000 | 131,072 |
qwen3-7-plus | qwen3.7-plus | 1,000,000 | 131,072 |
qwen3-6-flash | qwen3.6-flash | 1,000,000 | 65,536 |
qwen-plus-2025-07-28 | qwen-plus-2025-07-28 | 1,000,000 | 32,768 |
qwen-2-5-72b-instruct | qwen-2.5-72b-instruct | 32,768 | 16,384 |
qwen-2-5-7b-instruct | qwen-2.5-7b-instruct | 32,768 | 29,491 |
qwen-2-5-coder-32b-instruct | qwen-2.5-coder-32b-instruct | 32,768 | 29,491 |
qwen-plus-2025-07-28-thinking | qwen-plus-2025-07-28:thinking | 1,000,000 | 32,768 |
qwen3-14b | qwen3-14b | 131,072 | 16,384 |
qwen3-235b-a22b | qwen3-235b-a22b | 131,072 | 8,192 |
qwen3-235b-a22b-2507 | qwen3-235b-a22b-2507 | 262,144 | 16,384 |
qwen3-235b-a22b-thinking-2507 | qwen3-235b-a22b-thinking-2507 | 131,072 | 117,964 |
qwen3-30b-a3b | qwen3-30b-a3b | 131,072 | 16,384 |
qwen3-30b-a3b-instruct-2507 | qwen3-30b-a3b-instruct-2507 | 262,144 | 32,000 |
qwen3-30b-a3b-thinking-2507 | qwen3-30b-a3b-thinking-2507 | 81,920 | 32,768 |
qwen3-32b | qwen3-32b | 131,072 | 16,384 |
qwen3-5-122b-a10b | qwen3.5-122b-a10b | 262,144 | 81,920 |
qwen3-5-27b | qwen3.5-27b | 262,144 | 65,536 |
qwen3-5-35b-a3b | qwen3.5-35b-a3b | 262,144 | 16,384 |
qwen3-5-397b-a17b | qwen3.5-397b-a17b | 262,144 | 65,536 |
qwen3-5-9b | qwen3.5-9b | 262,144 | 235,929 |
qwen3-5-9b-batch | qwen3.5-9b:batch | 262,144 | 235,929 |
qwen3-5-flash-02-23 | qwen3.5-flash-02-23 | 1,000,000 | 65,536 |
qwen3-5-plus-02-15 | qwen3.5-plus-02-15 | 1,000,000 | 65,536 |
qwen3-5-plus-20260420 | qwen3.5-plus-20260420 | 1,000,000 | 65,536 |
qwen3-6-27b | qwen3.6-27b | 262,144 | 65,536 |
qwen3-6-35b-a3b | qwen3.6-35b-a3b | 262,144 | 235,929 |
qwen3-6-max-preview | qwen3.6-max-preview | 262,144 | 65,536 |
qwen3-6-plus | qwen3.6-plus | 1,000,000 | 65,536 |
qwen3-7-flash | qwen3.7-flash | 1,000,000 | 65,536 |
qwen3-8-2-4t-a95b | qwen3.8-2.4t-a95b | 1,048,576 | 262,144 |
qwen3-8-2-4t-a95b-batch | qwen3.8-2.4t-a95b:batch | 1,010,000 | 909,000 |
qwen3-8-27b | qwen3.8-27b | 1,000,000 | 131,072 |
qwen3-8-flash | qwen3.8-flash | 1,000,000 | 131,072 |
qwen3-8-max-0902 | qwen3.8-max-0902 | 1,000,000 | 131,072 |
qwen3-8b | qwen3-8b | 131,072 | 8,192 |
qwen3-coder | qwen3-coder | 262,144 | 65,536 |
qwen3-coder-30b-a3b-instruct | qwen3-coder-30b-a3b-instruct | 262,144 | 235,929 |
qwen3-coder-flash | qwen3-coder-flash | 1,000,000 | 65,536 |
qwen3-coder-next | qwen3-coder-next | 262,144 | 235,929 |
qwen3-coder-plus | qwen3-coder-plus | 1,000,000 | 65,536 |
qwen3-max | qwen3-max | 262,144 | 65,536 |
qwen3-max-thinking | qwen3-max-thinking | 262,144 | 65,536 |
qwen3-next-80b-a3b-instruct | qwen3-next-80b-a3b-instruct | 262,144 | 235,929 |
qwen3-next-80b-a3b-thinking | qwen3-next-80b-a3b-thinking | 262,144 | 235,929 |
The last four collapsed profiles are deprecated: they remain selectable so saved pipelines keep loading, but DashScope rejects their model IDs. They were introduced by OpenRouter fallback discovery in the model sync and carry OpenRouter/HuggingFace IDs rather than DashScope ones (e.g. DashScope uses qwen2.5-72b-instruct, not qwen-2.5-72b-instruct, and has no :thinking model variants — reasoning is controlled with the enable_thinking request parameter instead). Migrate to a live profile above.
Configuration
Choose a profile to set the Qwen model and token limits, then select the DashScope region that issued your API key. Profiles keep the model and token values fixed while exposing the API key, region, and model-source settings.
Authentication
Provide a DashScope API key in apikey. The key must start with sk-; anything else is rejected before any request is made. Make sure the key was issued for the region you select.
Notes
Regions
region selects the DashScope regional endpoint used for all API calls:
| Value | Region | Endpoint |
|---|---|---|
us | US (Virginia) | https://dashscope-us.aliyuncs.com/compatible-mode/v1 |
intl | Singapore | https://dashscope-intl.aliyuncs.com/compatible-mode/v1 |
cn | China (Beijing) | https://dashscope.aliyuncs.com/compatible-mode/v1 |
The default is us. An unrecognised value falls back to the US endpoint. DashScope API keys are region-specific, so a key issued for one endpoint will fail against another.
Setting base_url overrides this table entirely, which is how you reach a DashScope host Alibaba Cloud serves outside the three above — another Alibaba Cloud region, for instance. It is available on every live profile, including custom; the deprecated profiles above do not expose it. Leave it empty to use the regional endpoint.
Error handling
Provider exceptions are mapped to friendly messages instead of raw stack traces:
- Authentication failures surface as "Invalid DashScope API key."
- Rate-limit errors surface as "Rate limit exceeded. Please try again later."
- Connection failures surface as "Failed to connect to the DashScope API."
- Other DashScope API errors surface as "An error occurred with the DashScope API."
Rate-limit and connection errors are classified as retryable by the shared chat base; authentication and generic API errors are not retried.
Keeping the model list current
Profiles are maintained by the model sync tool, see tools/sync_models:
python tools/sync_models/src/sync_models.py --provider llm_qwen --enable-discovery --apply
Discovery — adding profiles — requires ROCKETRIDE_QWEN_KEY. Without it the command above still runs, but only enriches profiles that already exist: OpenRouter and LiteLLM can supply token counts, and neither may add a profile unless you also pass --allow-fallback-discovery. Avoid that flag here — it lets OpenRouter contribute the HuggingFace-style IDs DashScope does not accept, which is what the deprecated profiles above are. The stable aliases are listed in protected_profiles so a non-authoritative source cannot deprecate them.
Upstream docs
Schema
| Field | Type | Description | Default |
|---|---|---|---|
model | string | Model Qwen model | |
modelTotalTokens | number | Tokens Maximum context length in tokens | |
qwen.base_url | string | Base URL override Optional. Overrides the endpoint selected by Region. Leave empty to use the regional endpoint. Set this to reach a DashScope host other than the three listed above, for example another Alibaba Cloud region. | "" |
qwen.profile | string | Model Qwen AI model selection | "qwen-flash" |
qwen.region | string | Region DashScope regional endpoint. API keys are not interchangeable between regions. | "us" |
Dependencies
openailangchain-openailangchain-corelangchain