Skip to content

LLM Node Types & Workflow Orchestration Overview

Reference, synced 2026-06-13.

flowchart TD
  n0["Models"]
  n1["Overview"]
  n2["Products"]
  n3["Solutions"]
  n4["Pricing"]
  n5["Resources"]
  n6["Partners"]
  n7["Support"]
  n8["Language"]
  n0 --> n1
  n1 --> n2
  n2 --> n3
  n3 --> n4
  n4 --> n5
  n5 --> n6
  n6 --> n7
  n7 --> n8

Free access Accelerate Delivery with Fixed-Cost Agentic CodingWatch how it works

Models

Empowering AI innovation for both enterprises and developers with Alibaba Cloud’s best-in-class Qwen models, AI-native apps, and AI solutions.

Alibaba Cloud Model Studio \ Enterprise-grade large model service and application development platform.

Try Visual Model \ Supports image understanding, image generation, and video generation.

Models

HappyHorse-1.0-T2V \ Cinematic creative generation, ultimate dynamic details Qwen3-VL-Plus \ Native VL, spatial reasoning, 1M-context video analysis Wan2.7-VideoEdit \ Supports both localized and global editing with prompt

Qwen3.6-Plus \ Native multimodal, 1M context, agentic coding Wan2.7-Image-Pro \ Interactive editing, long-text rendering, precise prompt following Qwen-Plus \ Balanced intelligence, efficient inference, production-ready performance

Qwen-Image-2.0 \ Professional infographics, exquisite photorealism Z-Image-Turbo \ Ultra-fast image generation, high throughput, cost-optimized inference Qwen3-Coder-Next \ Multi-turn tool interactions, future-ready development support

Wan2.7-T2V \ High-fidelity T2V, 15s duration, advanced camera control Wan2.7-I2V \ Cinematic I2V with emotional depth and visceral impact Wan2.7-R2V \ Up to 5 mixed image/video inputs and audio timbre cloning

GenAI Application

Qoder \ Intelligent coding assistant, available for enterprise-dedicated deployment. Qoder CN \ AI-powered coding assistant that boosts developer productivity with intelligent code completion, AI chat, multi-file editing, and task automation.

AI Service

Model Experience \ Experience full-scale, multimodal model capabilities online. Platform for AI \ An AI-native algorithm engineering platform for end-to-end modeling, training, and inference service deployment. Fine-tune Video Generation Model \ Customize Wan’s text-to-video capabilities through model fine-tuning to meet your unique requirements.

AI Use Case

AI Savings Plan Hot \ Save up to 47% on AI costs. Limited-time offer tailored to your usage. AI Video Creation \ Elevate your professional video production with Wan 2.6.

AI Token Plan \ One plan. Multiple models. Big Savings with a Fixed Subscription. AI Image Creation \ All-in-one creative suite for copywriting, image generation, and poster design.

Overview

As a global full-stack AI leader, Alibaba Cloud aims to make computing accessible to everyone and help worldwide customers accelerate innovation.

Why Alibaba Cloud

About Alibaba Cloud \ AI Powered Cloud Technology Our Global Network \ Explore our global presence and deployment regions around the world Our Global Offices \ With offices in 4 continents, we're always close to where it matters.

Asia Accelerator \ Accelerate Success in Asia with Alibaba Cloud Go Global \ Benefits of our Global Alliance Trust Center \ Empowering enterprises with a secure, compliant, and globally trusted cloud infrastructure

Customers and Insights

Olympic Games \ Alibaba Cloud Powers Olympic Games with AI-powered cloud technology Case Studies \ Learn how customers are scaling their businesses on Alibaba Cloud Analyst Reports \ Learn what the top industry analyst firms are saying about Alibaba Cloud

What's New

Events and Webinars \ Quick access to upcoming and on-demand events Product Updates \ Stay informed of the latest innovations Press Room \ Latest news and media releases

Products

Featured ProductsAI & Machine Learning Computing Container Storage Networking & CDN Security Middleware Database Analytics ComputingMedia ServicesEnterprise Services & Cloud CommunicationDomain Names and WebsitesEnd User ComputingServerlessDeveloper ToolsMigration & O&M ManagementApsara Stack

Alibaba Cloud Model Studio \ Supercharge your AI journey effortlessly with industry-leading GenAI models ApsaraDB RDS \ Store and manage your business data, with automated monitoring and backups Certificate Management Service (Original SSL Certificate) \ Create a safe and secure connection between your website and users

Elastic Compute Service (ECS) \ Host websites anywhere and scale enterprise workloads Container Service for Kubernetes (ACK) \ Run and scale containerized applications on managed Kubernetes infrastructure Object Storage Service (OSS) \ Store large amounts of data in the cloud and access it anywhere, anytime

Simple Application Server (SAS) \ All-in-one services for fast deployment Elastic IP Address (EIP) \ Manage your public IPs independently to improve internet network quality Domain Names and Website \ Get the perfect domain name to suit your every need

Solutions

Solutions by Industry Technical Solutions AI WebsitesNetworking Security and ComplianceData and AnalyticsEnterprise Service and ApplicationCloud MigrationCloud NativeHybrid CloudSMB solutions

Financial Services \ Innovate faster with Alibaba Cloud Games \ Grow your game rapidly with high global availability

New Retail \ Alibaba Cloud enables digital retail transformation to fuel growth and realize an omnichannel customer experience throughout the consumer journey. Media and Entertainment \ Ready your content for today's media market with a digitalized media journey

Supply Chain \ Power your supply chain with intelligent, efficient, and reliable solutions Sports \ Digitizing the sports industry with intelligent tech

Sustainability \ Achieve a sustainable future with low-carbon and energy-efficient technologies

Pricing

Flexible options like pay-as-you-go and clear billing rules to meet diverse business needs.

Overview & Tools

Pricing Calculator \ Get an instant pricing estimate based on your usage and needs Free Trial \ Try our 80+ cloud products for free.

Pricing Options \ Get the most out of Alibaba Cloud with flexible pricing

Optimize your cost

Migrate & Save \ Superior Performance At Lower Pricing. Save up to 50%. Promotion Center \ Unlock the latest Alibaba Cloud offers & promos

Resources

Official documentation, extensive tools, training resources, and a community to grow and innovate in the cloud.

Technical Resources

Documentation \ Product guides and FAQs Architecture Center \ Design reliable, secure, and efficient cloud architecture. Intelligent Solution Explorer \ Find the right solution for you, powered by AI

Blog \ Latest cloud insights and developer trends Whitepapers \ Research that explores the how and why behind our technology

Training&Certification

Alibaba Cloud Academy \ Build cloud skills and earn certifications with expert-led training.

Developer Hub

Alibaba Cloud Project Hub \ Explore real-world projects built by developers using our platform. Our Developer MVPs \ Celebrating the developers who lead, build, and inspire our community

Partners

Partner-first strategy offering collaborative product, sales, and service models, plus high-quality partner solutions that complement Alibaba Cloud’s capabilities.

Marketplace

AI Alliance for ISVs \ Partner with us to build and grow AI solutions together ISV Benefits \ Unlock resources, market access, and go-to-market support as an ISV partner

Alibaba Cloud Marketplace \ Explore ready-to-deploy solutions from our partners and ISVs

Find a Partner

Partner Hub \ Find your ideal partner in no time

Become a Partner

Partner Network \ A partner portal for Alibaba Cloud Channel, Technology, MSP partner and other partner programs

Support

Full-lifecycle support and expert services, from cloud advisory and migration to operations.

Support & Professional Services

Professional Services \ Expert-led services to design, migrate, and optimize your cloud journey Support Plans \ Flexible support for every stage — from startup to enterprise

Partner Support Program \ Priority technical support for partners, with dedicated managers and faster issue resolution

Contact us

Connect With Us \

Talk to a sales expert and get a custom quote for your business

Language

  • English
  • 简体中文
  • 繁體中文
  • 日本語
  • Bahasa Indonesia

Locale

Visit aliyun.com

Documentation

Alibaba Cloud Model Studio

User Guide (Models) User Guide (Application) API Reference (Models) API Reference (Application)

Search for Help Content

Getting Started

The Beginner's Guide

Well-Architected Framework

AI & Machine Learning

Platform For AI

Alibaba Cloud Model Studio

DashVector

Artificial Intelligence Recommendation

OpenSearch

Image Search

Machine Translation

Intelligent Speech Interaction

Optimization Solver

Intelligent Computing LINGJUN

Computing

Elastic Compute Service

Elastic GPU Service

Elastic Container Instance

Dedicated Host

Compute Nest

Simple Application Server

Cloud Box

Auto Scaling

Elastic High Performance Computing

Batch Compute (Deprecated)

Function Compute

Serverless App Engine

ENS

Elastic Desktop Service

App Streaming

WUYING Terminal

Cloud Phone

Edge Network Acceleration

Alibaba Cloud Linux

AgentBay

Container

Container Service for Kubernetes

Container Compute Service

Container Registry

Storage

Object Storage Service

Cloud Parallel File Storage

File Storage NAS

Tablestore

Storage Capacity Unit

Simple Log Service

Cloud Backup

Intelligent Media Management

Drive and Photo Service

Data Transport

Cloud Storage Gateway

Data Online Migration

Hybrid Cloud Storage Array

Storage Services Overview

Backup and Disaster Recovery Center

Networking and CDN

Server Load Balancer

Elastic IP Address

Internet Shared Bandwidth

Data Transfer Plan

Virtual Private Cloud

NAT Gateway

PrivateLink

Alibaba Cloud DNS PrivateZone

Network Intelligence Service

Cloud Data Transfer

IPv6 Gateway

Anycast Elastic IP Address

Cloud Enterprise Network

Global Accelerator

VPN Gateway

Smart Access Gateway

Express Connect

CDN

Edge Security Acceleration

Cloud Network Well-architected Design Guidelines

Security

Anti-DDoS

Web Application Firewall

Cloud Firewall

Security Center

Bastionhost

Secure Access Service Edge

Certificate Management Service

Key Management Service

Data Security Center

Identity as a Service

Fraud Detection

AI Guardrails

Captcha

Blockchain as a Service

ID Verification

Managed Security Service

Middleware

Enterprise Distributed Application Service

Microservices Engine

Alibaba Cloud Service Mesh

SchedulerX

ApsaraMQ for RocketMQ

ApsaraMQ for Kafka

ApsaraMQ for RabbitMQ

ApsaraMQ for MQTT

Simple Message Queue (formerly MNS)

CloudFlow

EventBridge

Application Real-Time Monitoring Service

Managed Service for Prometheus

Managed Service for Grafana

Managed Service for OpenTelemetry

Performance Testing

STAROps

Databases

ApsaraDB Console

PolarDB

ApsaraDB RDS

ApsaraDB for OceanBase (Deprecated)

Tair (Redis® OSS-Compatible)

Lindorm

Time Series Database

ApsaraDB for MongoDB

ApsaraDB for HBase

ApsaraDB for Memcache

ApsaraDB for MyBase

AnalyticDB

ApsaraDB for ClickHouse

ApsaraDB for SelectDB

Data Transmission Service

Database Autonomy Service

Data Management

Database Gateway - Deprecated

ApsaraDB for Cassandra - Deprecated

Analytics Computing

MaxCompute

Hologres

Realtime Compute for Apache Flink

Elasticsearch

Vector Retrieval Service for Milvus

E-MapReduce

Data Lake Formation

DataV

Quick BI

Quick Audience

Quick Tracking

DataWorks

DataHub

Dataphin

Media Services

ApsaraVideo VOD

ApsaraVideo Live

Intelligent Media Services

ApsaraVideo Media Processing

Apsara Video SDK

Enterprise Services & Cloud Communication

Energy Expert

CloudQuotation

Salesforce on Alibaba Cloud

GoChina ICP Filing Assistant

Marketplace

Alibaba Mail

Direct Mail

Short Message Service

Voice Service

Phone Number Verification Service

Cell Phone Number Service

Chat App Message Service

Financial Intelligence Engine

Domain Names and Websites

Domain Names

ICP Filing

Alibaba Cloud DNS

End User Computing

Elastic Desktop Service

App Streaming

WUYING Terminal

Cloud Phone

AgentBay

Internet of Things

IoT Platform

Serverless

Serverless App Engine

CloudFlow

EventBridge

Simple Message Queue (formerly MNS)

Function Compute

Developer Tools

OpenAPI Explorer

Alibaba Cloud SDK

Cloud Shell

Resource Orchestration Service

Alibaba Cloud CLI

BSS OpenAPI

Terraform

Pulumi

Ticket System API

Mobile Platform as a Service

Alibaba Cloud DevOps

API Gateway

Cloud Control API

AI Coding Assistant Lingma

Cloud Skills Portal

Migration & O&M Management

CloudOps Orchestration Service

Cloud Monitor

Intelligent Advisor

Cloud Governance Center

ActionTrail

Cloud Config

Resource Access Management

Resource Management

Cloud Architect Design Tools

Migration Hub

Server Migration Center

Service Catalog

Logic Composer

Quota Center

CloudSSO

HTTPDNS

Solutions

SAP

SuperApp

OpenLake

Membership Service

Expenses and Costs

Account Center

More

Support

Legal

Tech Share Terms and Conditions

After Sales Support

China Gateway Program

Service Level Objectives

Management Console

Security Control

A workflow breaks down complex tasks into ordered steps to reduce system complexity. In Alibaba Cloud Model Studio, workflows let you combine nodes, such as large models, APIs, and Function Compute, to reduce coding costs.

Important

  • Console restrictions: Only International Edition users who created applications before April 21, 2025 can access the Application Development tab, as shown in the following figure.

This tab contains the following features: Applications ( agent application and workflow application), Components ( prompt engineering and plug-in), and Data ( knowledge base and application data). These are all preview features. Use them with caution in production environments.

  • API call limits: Only International Edition users who created applications before April 21, 2025, can call the application data, knowledge base, and prompt engineering APIs.

Application

Why use a workflow application

A workflow breaks down complex tasks into a sequence of steps to reduce system complexity. Creating a workflow application on Model Studio lets you define the execution order, assign responsibilities, and specify dependencies between steps to automate and optimize the process.

Common use cases for workflow applications include:

  • Travel planning: You can use a workflow plugin to select parameters, such as a destination, and automatically generate a travel plan that includes flights, accommodations, and attraction recommendations.

  • Report analysis: For complex datasets, you can combine data processing, analysis, and visualization plugins to generate structured, formatted analysis reports.

  • Customer support: You can use automated workflows to handle customer inquiries, including issue classification, to improve response speed and accuracy.

  • Content creation: You can generate content such as articles and marketing copy. Provide a topic and requirements, and the system automatically generates a draft.

  • Education and training: You can use a workflow to design personalized learning plans with progress tracking and assessments, enabling self-paced learning for students.

  • Medical consultation: Based on patient-entered symptoms, a workflow can combine multiple analytical tools to generate a preliminary assessment or recommend relevant tests to assist doctors with further diagnosis.

Examples

Example 1: Detect scam messages

This example demonstrates how to create a workflow application to determine if a text message is a potential scam. The workflow uses a start node, a large model node, and an end node.

1. Go to the App Management page. Click Create Application > Workflow Application, enter an App Name, and then click Create Now.
2. Add a large model node to detect scam messages. Drag the LLM node from the left-side pane to the canvas and configure the following parameters in its configuration panel. Leave other parameters at their default values. 1. Model Configuration: Select Qwen-Plus-latest. 2. Prompts: plaintext Analyze the provided message to determine if it is a potential scam. Provide a definitive answer. Requirements: Carefully review the message content, focusing on keywords and typical scam patterns, such as requests for urgent transfers, providing personal information, or promising unrealistic benefits. Steps: 1. Identify key elements in the message, including but not limited to the sender's identity, the request made, promised rewards, and any sense of urgency. 2. Compare the message against the characteristics of known scams to check for similar tactics or language patterns. 3. Evaluate the overall reasonableness of the message, considering whether the request aligns with standard procedures and common sense. 4. If the message contains links or attachments, do not click or download them to avoid potential security risks, and warn the user about the dangers of such content. Output Format: Clearly state whether the message exhibits characteristics of a scam and briefly explain your reasoning. If a scam is suspected, provide suggestions or preventive measures to protect the user. 3. User Prompt: plaintext Determine if the message "${sys.query}" is a potential scam. Close the panel and connect the Start node to the Large Model node.
3. Connect the large model node to the end node. Then, click the end node and in its configuration panel's editor, enter / to insert the Large Model 1/result variable. Leave other parameters at their default values.Configuration example:
4. Click Test in the upper-right corner. Enter Your package has been at the pickup station for several days. Please come and collect it. and click . After the workflow finishes running, view the output.
5. Next, enter You have a message saying you've won 1 million. Please check it. and click . After the workflow runs, view the output.

Example 2: Smart shopping assistant

This example shows how to create a smart shopping assistant using a workflow that helps users select mobile phones, TVs, and refrigerators. The workflow uses a start node, an intent classification node, large model nodes, and an end node.

1. Go to the App Management page. Click Create Application > Workflow Application, enter an App Name, and then click Create Now.
2. Drag the Intent Classification node from the left-side pane to the canvas and configure the following parameters in its configuration panel. Leave other parameters at their default values. - Input Variable: Select Built-in Variable > query. - Select Model: Select Qwen-Plus-latest. - Intent Classification: Click Add Intent and add the following three intents: - Buy TV - Buy mobile phone - Buy refrigerator - Memory: Enables the model to remember previous chat content during a conversation. Turn on the toggle and select Custom Cache. > Current Node Cache: The model remembers only the chat content within the current node. > Custom Cache: The model remembers the conversation across the entire workflow. Close the panel and connect the Start node to the Intent Classification node.
3. Drag a LLM node from the left-side pane to the canvas. In its configuration panel, configure the following parameters, leaving others at their default values. Then, close the panel and connect the Buy refrigerator output of the Intent Classification node to the Large Model node. - Model Configuration: Select Qwen-Plus-latest. - Prompts: ```plaintext You are a smart shopping assistant responsible for recommending refrigerators to customers. You must proactively ask the user for refrigerator parameters in the order listed in the [Refrigerator Parameter List] below. Ask for only one parameter at a time, and do not repeat questions for the same parameter. If the user provides a value for a parameter, continue to ask about the remaining parameters. If the user asks for the definition of a parameter, provide an explanation based on your professional knowledge, and then continue to ask which parameter value they prefer. If the user indicates they no longer wish to make a purchase, output: Thank you for visiting. We look forward to serving you next time. [Refrigerator Parameter List] 1. Usage Scenario: [Home, Small commercial, Large commercial] 2. Capacity: [200L, 300L, 400L, 500L] 3. Energy Efficiency Rating: [Level 1, Level 2, Level 3] After all parameters from the [Refrigerator Parameter List] have been collected, ask: "Are you sure you want to purchase?" while also showing the parameters the customer has selected, for example: For small commercial use300L
4. Drag a LLM node from the left-side pane to the canvas. In its configuration panel, configure the following parameters, leaving others at their default values. Then, close the panel and connect the Buy TV output of the Intent Classification node to the Large Model node. - Model Configuration: Select Qwen-Plus-latest. - Prompts: ```plaintext You are a smart shopping assistant responsible for recommending TVs to customers. You must proactively ask the user for TV parameters in the order listed in the [TV Parameter List] below. Ask for only one parameter at a time, and do not repeat questions for the same parameter. If the user provides a value for a parameter, continue to ask about the remaining parameters. If the user asks for the definition of a parameter, provide an explanation based on your professional knowledge, and then continue to ask which parameter value they prefer. If the user indicates they no longer wish to make a purchase, output: Thank you for visiting. We look forward to serving you next time. [TV Parameter List] 1. Screen Size: [50-inch, 70-inch, 80-inch] 2. Refresh Rate: [60Hz, 120Hz, 240Hz] 3. Resolution: [1080P, 2K, 4K] After all parameters from the [TV Parameter List] have been collected, ask: "Are you sure you want to purchase?" while also showing the parameters the customer has selected, for example: 50-inch120Hz
5. Drag a LLM node from the left-side pane to the canvas. In its configuration panel, configure the following parameters, leaving others at their default values. Then, close the panel and connect the Buy mobile phone output of the Intent Classification node to the Large Model node. - Model Configuration: Select Qwen-Plus-latest. - Prompts: ```plaintext You are a smart shopping assistant responsible for recommending mobile phones to customers. You must proactively ask the user for mobile phone parameters in the order listed in the [Mobile Phone Parameter List] below. Ask for only one parameter at a time, and do not repeat questions for the same parameter. If the user provides a value for a parameter, continue to ask about the remaining parameters. If the user asks for the definition of a parameter, provide an explanation based on your professional knowledge, and then continue to ask which parameter value they prefer. If the user indicates they no longer wish to make a purchase, output: Thank you for visiting. We look forward to serving you next time. [Mobile Phone Parameter List] 1. Usage Scenario: [Gaming, Photography, Watching movies] 2. Screen Size: [6.4-inch, 6.6-inch, 6.8-inch, 7.9-inch foldable screen] 3. RAM + Storage Space: [8GB+128GB, 8GB+256GB, 12GB+128GB, 12GB+256GB] After all parameters from the [Mobile Phone Parameter List] have been collected, ask: "Are you sure you want to purchase?" while also showing the parameters the customer has selected, for example: For photography8GB+128GB
6. Drag a Variable Processing node from the left-side pane to the canvas. In its configuration panel, configure the following parameters, leaving others at their default values. Then, close the panel and connect the Default intent output of the Intent Classification node to the Variable Processing node. - Output Mode: Select Text Output. - In the editor box, enter: plaintext Hello! Please tell me if you want to buy a TV, a refrigerator, or a mobile phone.
7. Connect all three large model nodes and the Variable Processing node to the End node. Then, click the end node and, in its configuration panel's editor, enter / to insert the following four variables: Large Model 1/result, Large Model 2/result, Large Model 3/result, and Variable Processing 1/result. Leave other parameters at their default values.Configuration example:
1. Click Test in the upper-right corner. In the dialog box, enter Tell me about your refrigerators, click , and view the output. 2. Next, enter Home use and view the output. 3. Then, enter 200L and view the output. 4. Finally, enter Level 1 energy efficiency and view the output.

Session parameters

A session variable acts as a global variable, storing parameters throughout the workflow's lifecycle to be referenced by any node.

Click the icon in the upper-right corner of the canvas configuration page.

Node

Nodes are the core functional units of a workflow application. Each node performs a specific task, such as executing an action, triggering a condition, processing data, or directing the workflow. Combine these nodes like building blocks to create automated processes.

Start and end

Large model

Knowledge base

API

Plugin

Function Compute

Script

Condition

Intent classification

Flow output

Variable processing

Parameter extraction

Agent group

Create agent

Multimodal generation

  • When to use

    • When designing a workflow, define the structure and content of the input/output parameters in the start and end nodes.
  • How to use

    • Start node

      ComponentDescription
      Predefined variablesThe workflow provides the following predefined variables to process user input and maintain conversation history: - query: Stores the user's text input. - historyList: Stores conversation history to maintain context in a multi-turn conversation. To use this variable in nodes that support the memory feature (such as large model and intent classification nodes), select Custom Cache. - imageList: Stores user-uploaded images to enable image analysis or multimodal conversation. To use this variable in nodes that support the memory feature (such as large model and intent classification nodes), select Custom Cache.
      Custom variablesCustom variables are structured input parameters that you define for a workflow. They receive data from tests or API calls and can be referenced in subsequent nodes. When creating a custom variable, configure the following parameters: - Variable name: Enter a meaningful name. Chinese characters are not supported. - Type: The data type of the variable. Supported types are String, Boolean, Number, Object, Array (of String, Boolean, Number, or Object), and File. - Description: Briefly describe what the variable does and when to use it.
    • End node

      ComponentDescription
      Output modeText Output: Suitable for unstructured content. In the input box, you can enter fixed content or type / to reference variables, which determines the final result returned to the user. You can source variables from the output of any workflow node or from session variables. Properly mapping output variables allows you to control the workflow's data flow and ensure an accurate and complete final response. JSON Output: Suitable for outputting structured content in JSON format. You can define variable names and enter text or reference variables.
      Streaming outputThe Streaming Output switch applies only to Text Output mode. When enabled, responses from large model and application component nodes are streamed token by token. When disabled, a full response is returned after it is completely generated.
  • Why use it

This node is the intelligent core of your workflow. It can understand language, generate text, analyze images, and engage in multi-turn conversations. You can use it to write copy, summarize text, or even analyze image content (using Qwen-VL models).

  • Features

    • Supports both single-input processing and batch processing of large datasets.

    • You can configure different large models, such as Qwen-Plus, so you can choose the right model based on performance, latency, or other requirements.

  • Node parameter configuration

ParameterDescription
Mode selectionSingle Mode: Calls the large model once.
Batch Mode: The node runs multiple times, sequentially assigning an item from a list to a batch variable in each run. This process stops when all list items are processed or the configured maximum batch runs is reached. Batch Configuration: - Maximum Number of Batches (Range: 1–100; default for standard users: 100): The maximum number of times the batch process will run. Note The number of batch runs depends on the minimum length of the input arrays. If no input variable is configured, the number of runs is determined by the configured batch count. - Number of Parallel Runs (Range: 1–10): Sets the concurrency limit for batch processing. A setting of 1 executes all tasks serially.
Model configurationSelect a suitable large model and adjust its parameters. The available models are listed on the UI. For a detailed introduction to the models, see the model list. For information on the API call rate limits for each model, see rate limiting.

| Parameter configuration | Temperature: Controls the diversity of the generated content. A higher temperature value increases the randomness of the generated text, producing more unique outputs, while a lower value makes the content more conservative and consistent. > This configuration is not currently supported for DeepSeek R1 series models. | | Maximum Reply Length: Limits the maximum length (in tokens) of the text generated by the model, not including the prompt. This limit varies by model. | | The following parameters are available only when a Qwen-VL model is selected: - Model Input Parameters: For vlImageUrl, you can reference a variable or input an image URL. - Image Source: - Image set: The model treats each uploaded image independently, selecting the most relevant one to answer the user's question. A single image can be passed directly. For example: https://****.com/****.jpg. Multiple images can be passed as a list. For example: ["URL","URL","URL"]. - Video frames: The model considers the uploaded images to be from the same video and understands them sequentially as a whole. You must provide a minimum of 4 video frames. | | System prompt | Provides system-level instructions to the model. You can use it to define the model's role, task, or output format. For example: "You are a math expert who specializes in solving math problems. Please output the problem-solving process and result in the correct format." | | User prompt | The user's input to the model, such as a request or instruction. You can configure a prompt template or insert variables. The model uses this input to generate a response. | | Memory | Memory is the context the model retains across multiple turns in a conversation. - Current node cache: Uses the output of this node as context. The model will only remember context generated within this node. - memory rounds: The number of conversation rounds to remember. One input and its corresponding output constitute one round. - Custom cache: Uses a specified context variable as the context. - Context variable: Select the source of the context. | | Output variable | The name of the variable that stores the processing result of this node. Subsequent nodes can use this variable to access this node's result. > DeepSeek R1 series models support outputting the Deep Thinking process (reasoningContent). | | Retry on failure | If disabled, the node stops executing when an error occurs. If enabled, the node re-executes on error, according to the configured number of retries and interval. - maximum retries: The maximum number of retry attempts when a request fails. - retry interval: The time interval between each retry attempt, in milliseconds. | | Exception handling | If disabled, the node handles errors using the system's default error handling mechanism. If enabled, the node executes custom processing logic according to the configuration when an error occurs. - Default value: The content in result is output when an exception occurs. - exception branch: If an exception occurs, this branch is executed. You must configure its processing flow. | | Return results | This feature is deprecated. This option applies only to API calls and determines whether the node's content is output. To learn more about the purpose of this component, see Application calling. |

Note

To integrate applications via API, see Application calling.

  • Definition

The knowledge base node implements retrieval-augmented generation (RAG). It retrieves relevant chunks from a specified knowledge base based on the input, and passes the results as context to a downstream large model node. This process addresses common problems with large models, such as outdated knowledge, inability to access private data, and hallucinations.

  • Input configuration

Defines the criteria for querying the knowledge base.

ParameterDescription
contentThe text used to query the knowledge base. You can enter text directly or type / in the editor to select and reference an output variable from an upstream node.
imageListThe images used to query the knowledge base. You can enter a publicly accessible image URL, such as https://xxx.xxx.com/xxx/xxx.jpeg, or type / in the editor to select and reference an output variable from an upstream node.
  • Knowledge base selection method
Selection methodApplicable scenariosConfiguration
Select Fixed Knowledge BaseEach call uses the same, pre-selected knowledge base.Select a specific knowledge base from the drop-down list. You can add three types of knowledge bases: documents, tables, and images.
Dynamic ImportThe output of an upstream node dynamically determines which knowledge bases to use.Configure the CodeList variable. You can specify a list of knowledge bases by referencing a variable or by direct input.
  • Knowledge base invocation method

This setting defines the trigger logic for the knowledge base. The default method is always invoke.

Invocation methodDescriptionParameters
always invokePerforms a knowledge base retrieval for each user input. This method is suitable for high-frequency question-and-answer scenarios.None
intelligent invokeThe agent determines whether to query the knowledge base based on the user's input and its description. This is suitable for flexible conversational scenarios.Knowledge Base Description (Required)
legacy invocationResults from document, table, and image knowledge bases are recalled together. A single topK parameter controls the number of recalled chunks, and does not support separate debugging for each type.topK (Range: 1–50, Default: 10)

Parameter descriptions:

  • Knowledge base description: Describes the content of the knowledge base and the conditions under which its data can be used for a response. Supports text input or entering / to insert a variable.

  • topK: The maximum number of chunks to recall from each knowledge base. The actual number of recalled chunks depends on matching results and may be less than the topK value.

  • Debug recall results

This feature is supported only in always invoke and intelligent invoke modes. Click the Debug button next to a document, table, or image type to test the recall results for that knowledge base. Use the results to fine-tune your configuration. When you are done, click Save in the upper-right corner to apply the configuration to the current application.

  • Knowledge base filtering

This feature is disabled by default. When enabled, the system uses a large model to intelligently filter the recalled results from documents and tables, improving the final output quality. This feature is suitable for scenarios that require high answer precision or need to filter out irrelevant information.

  • Node output

The knowledge base node outputs a structured object containing a result object:

FieldTypeDescription
resultObjectContains all information returned from the current query.
result.chunkListArray<Object>An array of recalled chunks. This array is empty if the query returns no results.
result.chunkList[ ].contentStringThe original content of the recalled knowledge base chunk.
result.chunkList[ ].titleStringThe title of the document to which the chunk belongs.
result.chunkList[ ].documentNameStringThe name of the document where the recalled chunk is located.
result.chunkList[ ].scoreNumberThe similarity score of the knowledge base chunk. A higher score indicates a better match.
result.chunkList[ ].idStringThe chunk ID.
result.chunkList[ ].dataIdStringThe document ID.
result.chunkList[ ].docUrlStringThe source file download link.
result.chunkList[ ].knowledgeBaseIdStringThe knowledge base ID.
result.chunkList[ ].nidStringThe identifier for the original text.
result.chunkList[ ].imagesArray<String>A list of images.
result.chunkList[ ].pageNumberArray<Number>An array of page numbers.
result.rewriteQueryStringThe rewritten user query.

Note

To ensure the API node can access the target service, add the Model Studio service IP addresses (47.93.216.17, 39.105.109.77, 60.205.180.248, and 59.110.152.173) to the allowlist of the inbound rules in your target server's security group or firewall.

  • Definition

Calls a custom API service using the POST, GET, PUT, PATCH, or DELETE method and outputs the API call result.

  • Parameter configuration
ParameterDescription
API addressThe request URL for the API. - Request method: - POST: Submits data to the server to create a new resource. - GET: Retrieves a resource without modifying data on the server. - PUT: Updates a specified resource or creates a new resource on the server. - PATCH: Partially updates a resource on the server. - DELETE: Deletes a specified resource from the server. - URL: Enter the full API address in the input field. You can enter the address directly or use a variable to reference a value from previous nodes. For example, https://dashscope.aliyuncs.com/compatible-mode/v1/files.
Header settingsSpecifies the HTTP request headers, such as Content-Type and Authorization. You can enter values directly or reference them from previous nodes.
Parameter settingsQuery parameters in the request URL. You can enter values directly or reference them from previous nodes. For example, limit in https://dashscope.aliyuncs.com/compatible-mode/v1/files?limit=5.
Body settingsSelect the correct type based on the API's requirements. - none: No request body. Suitable for GET requests. - form-data: Form data, used for file uploads or key-value pairs. - raw: Raw text, such as JSON or XML. - JSON: An automatically formatted JSON object.
Timeout setting (s)The maximum time to wait for a response. If a response is not received within this time, the request times out.
Retry on failureIf disabled, the node stops executing when an error occurs. If enabled, the node retries the request according to the configured maximum retries and retry interval when an error occurs. - Maximum retries: The maximum number of times to retry a failed request. - Retry interval: The interval between retries, in milliseconds.
Exception handlingIf disabled, the node uses the default error handling mechanism when an error occurs. If enabled, the node executes custom logic when an error occurs. - Default value: Outputs the content of result when an exception occurs. - Exception branch: Proceeds to the exception branch if an exception occurs. You must configure the processing flow for this branch.
OutputStores the API response in a specified variable for use by subsequent nodes.

Note

For details on integrating applications with an API, see application integration.

  • Definition

You can add a plugin node to your workflow application to perform more complex tasks. Model Studio offers a range of official plugins, such as Quark Search, Calculator, and a Python code interpreter. You can also create custom plugins to meet specific needs.

See Plugin overview.

  • Prerequisites

    • To use an official or third-party plugin, you must first request its activation. For detailed instructions, see Plugin overview.

    • If you want to reference a custom plugin, you need to create a custom plugin first.

  • Parameter configuration

ParameterDescription
InputSpecifies the content for the node to process. The variables change depending on the selected plugin. When you select a plugin, its built-in variables load automatically. You can go to the Official plugins or Custom plugins page, click a plugin card, and view how to configure its variables.
OutputThe output variable that contains the node's processing result. Subsequent nodes use this variable to access the result.
Retry on failureIf disabled, the node stops executing when an error occurs. If enabled, the node retries execution based on the configured number of retries and interval when an error occurs. - Maximum retries: The maximum number of retry attempts when a request fails. - Retry interval: The time interval between each retry, in milliseconds.
Exception handlingIf disabled, the node uses the system's default error handling mechanism. If enabled, the node executes custom logic when an error occurs. - Default value: Outputs the content of result when an exception occurs. - Exception branch: Executes this branch if an exception occurs. You must configure the processing flow for the branch.
  • Definition

Calls a custom service on Alibaba Cloud Function Compute.

  • Parameter configuration

Important

The default timeout for a Function Compute node is 60 seconds and cannot be changed.

ParameterDescription
InputThe input variable that this node processes. You can reference variables from upstream nodes or enter values directly.
RegionSelect the region where your Function Compute service is located: Singapore, Kuala Lumpur, or Indonesia (Jakarta).
Service configurationSelect the Function Compute service to invoke. You must first create a function. The Alibaba Cloud account that creates the Function Compute service must be the same as the account you use to log in to Model Studio (Bailian), or it must belong to the same Alibaba Cloud root account.
OutputThe variable that stores the node's output. Subsequent nodes can reference this variable.
  • Definition

Uses a script to transform input into a specific template or output format. This involves parsing, transforming, and formatting the data for consistency and readability.

  • Parameter configuration
ParameterDescription
inputAccepts a fixed value, or a reference to a variable from an upstream node or a session variable.
codeDefines the core logic for the node. Both JavaScript and Python are supported. - Get input: Use the built-in params object to access input parameters in the format params.<variable_name>. For example, params.input1. - Return output: The main handler function must return a dictionary or object. The key-value pairs of this object become the node's output.
outputThe dictionary returned by the return statement in the code serves as the output of this node. For example, if you return {'result': 'Processing successful'}, downstream nodes can access the string "Processing successful" by using node_name.result.
retry on failureIf disabled, the node stops executing when an error occurs. If enabled, the node retries execution according to the configured maximum retries and retry interval. - maximum retries: The maximum number of retry attempts after a failure. - retry interval: The interval between retries, in milliseconds.
exception handlingIf disabled, the node uses the system's default error handling mechanism. If enabled, the node executes custom logic when an error occurs. - Default value: Outputs the content of result when an exception occurs. - exception branch: Routes execution to a separate, user-configured branch if an exception occurs.
  • Definition

The condition node adds branching logic to a workflow. When a variable meets a specified condition, the workflow follows the corresponding downstream link. The node supports AND/OR configurations and evaluates conditions sequentially from top to bottom.

  • Parameter configuration
ParameterDescription
conditional branchEnter the conditional expressions. Conditions in different groups are joined by OR, while conditions within the same group are joined by AND.
elseThe default output when no defined conditions are met.
  • Definition

Intelligently matches user input with a predefined intent based on its description, and then routes to the corresponding link.

  • Parameter configuration
ParameterDescription
input variableSpecifies the input variable to be classified. You can enter a value directly, or use a variable from a preceding node or a session variable.
model selectionSupports Qwen-Plus and the intent classification model.
intent classificationDefine the intents for the model to classify. The model matches the input to an intent based on its description and then routes to the corresponding link. For example: "For calculating math problems" or "For answering weather-related questions".
Other intentThis branch executes when the model determines that the input does not match any of the defined intents.
intent mode- single-select mode: The large model selects the single best-matching intent from the available configurations. - multi-select mode: The large model selects all matching intents from the available configurations.
reasoning mode- fast mode: This mode improves processing speed by avoiding complex reasoning steps, making it suitable for simple scenarios. - quality mode: This mode uses step-by-step reasoning to achieve more accurate classification.
memoryRefers to the model's ability to remember conversation history in multi-turn conversations, which serves as the context. - node-local cache: Uses the output of this node as context. The model's memory is limited to the context generated by this node. - memory turns: The number of conversation turns to remember. One input and its corresponding output constitute one turn. - custom cache: Uses the content of a specified context variable as the context. - context variable: Specifies the source of the context.
promptProvides additional instructions or constraints for the intent classification model. You can add examples or define constraints to help the model's classification results better match your requirements. Example Suppose you are developing a customer service system for an e-commerce platform where users might ask about order inquiries, returns, or payments. To ensure accurate classification, you can add relevant prompts and examples in the advanced settings. markdown Please classify the intent based on the following examples: Example 1: User input "I want to return the jacket I just bought", classify as "Returns and exchanges". Example 2: User input "Please help me check the shipping status of my order", classify as "Order inquiry". Constraints: Only process order-related inquiries. Ignore payment and technical issues. Effect: User input: "When will the book I ordered from your website last week be delivered to my home?" Classification result: "Order inquiry" In this example, providing specific classification examples guides the model to classify the "delivery time query" as an "Order inquiry" intent, while the constraints limit the classification scope to exclude other irrelevant questions.
outputThe variable that stores the node's output. Downstream nodes can use this variable to access the result.

Note

Running this node consumes tokens. The number of tokens consumed is displayed at runtime.

  • Definition

The flow output node outputs specified content at any point in a workflow. This differs from the final output, which is returned only after the workflow completes.

The flow output node is useful in the following scenarios:

  • Split large outputs: If a workflow generates a large amount of output, you can insert a flow output node to split the content into two parts. One part is output by the flow output node and the other by the end node.

  • Retrieve intermediate information: Insert this node at key points in a workflow to output information such as variable values or process status. This helps you track progress and troubleshoot issues.

  • Parameter configuration

ParameterDescription
Output contentEnter fixed content or type / to reference variables from preceding nodes or session variables.
Streaming outputWhen enabled, content from a large model is streamed to the conversation token by token. When disabled, the entire response is generated first and then output at once.
  • Definition

Used to transform and process text content, such as by extracting specific content or converting formats. This node also supports template mode.

  • Parameter configuration
ParameterDescription
Output mode- Text output: Converts the input to text. In the editor, you can either enter the content directly or insert variables from upstream nodes or session variables. - JSON output: Formats the input variables as a JSON object. - Group aggregation: Determines which values are returned from a group according to a specified strategy. You can return either the first or the last non-empty value from each group.
Output variableSpecifies the variable to store the node's output for use by subsequent nodes.
Return resultThis feature is being deprecated. Applicable only to API calls, this parameter determines whether the node's output is returned. See application call.
  • Definition

Extracts structured parameters from text using a model.

  • Parameter configuration
ParameterDescription
InputThe text to extract parameters from.
Model selectionSpecifies the model used to extract parameters. Choose the one that best suits your business needs.
Parameters to extractThe model extracts parameters from the input based on the name, type, and description you provide.
PromptAdditional instructions to guide the model during parameter extraction.
MemoryMemory provides conversational context, helping the model remember previous interactions in multi-turn scenarios. - Cache in this node: Uses the output of this node as context. The model only remembers context from this node. - Memory turns: The number of conversation turns to remember. One turn consists of one input and its corresponding output. - Custom cache: Uses a specified context variable to provide context. - Context variable: Select the source of the context.
OutputContains the parameters extracted by the model. Subsequent nodes can reference this output as a variable.
  • Definition

An agent group contains multiple sub-agents. Based on task requirements, the decision model within the group automatically plans the execution flow, schedules sub-agents, and coordinates them to complete the task. Use this node if you need to complete a large project but are unsure how to design the specific workflow. By breaking down a complex task into sub-tasks that different agents process in parallel, you can improve task execution speed.

  • Parameter configuration
ParameterDescription
InputSpecifies the content that the node will process. You can reference variables from a previous node.
Model selectionSelect a decision model.
Group nameEnter a custom name for the agent group.
AgentsSpecifies the sub-agents within the agent group. Only published agent applications can be added. After adding a sub-agent, click Configure to describe its function. This description helps the decision model determine which sub-agent to invoke for the current task.
Output variableSpecifies the variable for storing the output of this node. This variable can be referenced by subsequent nodes.
  • Definition

Creates a new agent exclusively for the orchestration canvas.

  • Parameter configuration
ParameterDescription
InputThe content to be processed by this node. You can reference variables from preceding nodes.
Agent nameA custom name for the agent.
Model selectionThe large model used by the agent.
PromptUse natural language to define the agent's role and tasks.
Knowledge baseSelect a knowledge base for the agent.
PluginEnables the agent to call official or custom plugins to extend its capabilities.
Output variableThe variable that stores the node's output for use in subsequent nodes.
  • Definition

The multimodal generation node uses Alibaba Cloud multimodal models to generate image, video, or audio content based on the configured prompts and parameters. Use this node for content creation, marketing asset production, short video generation, and voice-over synthesis.

  • Model selection

Model selection

Select a multimodal generation model from the drop-down list. Models are categorized by function:

Model typeSupported modelsFeatures
Image generationQwen-ImageExcels at rendering complex Chinese and English text.
Wan 2.6-image-generationGenerates and edits images from text instructions, supporting both image editing and mixed text-and-image modes.
Qwen-Image-PlusExcels at rendering complex Chinese and English text.
Z-Image-TurboA lightweight text-to-image model that quickly generates high-quality images. It supports bilingual rendering (Chinese and English), complex semantic understanding, and multiple styles and themes.
Wan 2.2-text-to-image-PlusThe Pro edition of Wan 2.2, with comprehensive improvements in creativity, stability, and photorealism.
Wan 2.2-text-to-image-FlashThe Flash edition of Wan 2.2, with comprehensive improvements in creativity, stability, and photorealism.
Video generationWan 2.6-text-to-videoWan 2.6. Features multi-shot storytelling capabilities and supports both automatic voice-overs and custom audio file inputs.
Wan 2.6-image-to-videoWan 2.6. Features multi-shot storytelling capabilities and supports both automatic voice-overs and custom audio file inputs.
Wan 2.5-text-to-video-PreviewWan 2.5 preview. Supports automatic voice-overs and custom audio file inputs.
Wan 2.5-image-to-video-PreviewWan 2.5 preview. Supports automatic voice-overs and custom audio file inputs.
Wan 2.2-text-to-video-PlusThe Pro edition of Wan 2.2. Offers more accurate instruction understanding, stable and smooth motion generation, and richer details in the output.
Wan 2.2-image-to-video-PlusThe Pro edition of Wan 2.2, offering more accurate instruction understanding, controllable camera movement, and consistent visuals. It also features comprehensive improvements in stability, success rate, and content richness.
Wan 2.2-image-to-video-FlashThe Flash edition of Wan 2.2, delivering ultra-fast generation, more accurate instruction understanding and camera control, consistent visuals, and comprehensive improvements in stability and success rate.
Audio generationQwen3-TTS-FlashSupports mixed-language text input and streaming audio output.

The list of available models is subject to change. Refer to the UI for the most current options. For model invocation pricing, see Model Invocation Billing.

  • Configuration parameters vary by model. The panel shows only the settings relevant to your selection.

Image generation

Video generation

Audio generation

After you select an image generation model, configure the following parameters. The available parameters vary by model; refer to the UI for the specific options.

ParameterRequiredDescription
Positive promptYesDescribes the content, style, composition, and other elements of the target image. You can enter text directly or type / to insert an output variable from an upstream node.
Negative promptNoDescribes content that you do not want in the image, such as blurriness or watermarks. You can enter text directly or type / to insert an output variable from an upstream node.
Size/ResolutionYesThe dimensions of the image.
prompt_extendNoIf enabled, the system uses a large model to intelligently rewrite the input prompt. This feature applies only to the positive prompt. It significantly improves results for shorter prompts but adds 3-4 seconds of latency.
Add WatermarkNoControls whether to add an Alibaba Cloud watermark to the generated image. This feature is disabled by default.
Reference imageNoUpload a reference image to guide the style or content of the generated image. You can pass a File-type variable from the start node or a public URL.
Number of ImagesNoThe number of images to generate in a single request. The default value is 1.
Intelligent prompt rewritingNoIf enabled, a large model intelligently rewrites the input prompt. This feature applies only to the positive prompt. It significantly improves results for shorter prompts but increases the image generation time.
enable_interleaveNoDisabled (default): Image editing mode. Edits, performs style transfer, or generates consistent subjects using 1 to 4 input images. Enabled: Mixed text-and-image output mode. Generates a single content block containing both text and images, based on an image or plain text.
Random seedNoControls randomness in the generated output (default: 1234). Using the same seed value and prompt produces similar images.
Intelligent ThinkingNoIf enabled, the large model performs reasoning and prompt rewriting. This can improve the generation quality but also increases the generation time.

After you select a video generation model, configure the following parameters. The available parameters vary by model; refer to the UI for the specific options.

ParameterRequiredDescription
Positive promptYesDescribes the scene, actions, style, and other elements of the target video. You can enter text directly or type / to insert an output variable from an upstream node.
Negative promptNoDescribes content that you do not want in the video, such as blurriness or watermarks. You can enter text directly or type / to insert an output variable from an upstream node.
ResolutionYesThe video resolution. The available options vary by model.
Video DurationYesThe duration of the generated video. The available options vary by model.
Random seedYesControls randomness in the generated output. Using the same seed value and prompt produces similar videos.
Intelligent ExpansionNoIf enabled, the system uses a large model to intelligently rewrite the input prompt to improve the video generation quality. This feature applies only to the positive prompt.
Smart Multi-shotNoIf enabled, the model generates a multi-shot video.
Generate AudioNoIf enabled, you can add a voice-over by providing a public audio file URL.
Reference imageYes (image-to-video models only)Specifies a reference image to serve as the video's first frame. The model generates the video based on this image. You can pass a File-type variable from the start node or a public URL.
ParameterRequiredDescription
Text for SynthesisYesThe text content to convert to speech. You can enter text directly or type / to insert an output variable from an upstream node.
Language TypeNoThe language for the synthesized speech (default: Chinese). Supported languages: Chinese, English, German, Italian, Portuguese, Spanish, Japanese, Korean, French, and Russian.
VoiceNoThe voice for the synthesized speech. For a list of supported voices, see Real-time Speech Synthesis - Qwen.
  • Node output

After the multimodal generation node runs, it outputs the following variables:

ParameterTypeDescriptionExample
outputObjectThe complete output object, containing the task status, execution time, and results.-
output.task_statusStringThe task status, such as "SUCCEEDED" or "FAILED".SUCCEEDED
output.submit_timeStringThe task submission time, in YYYY-MM-DD HH:mm:ss.SSS format.2026-01-22 10:54:44.200
output.end_timeStringThe task end time, in YYYY-MM-DD HH:mm:ss.SSS format.2026-01-22 10:54:51.685
output.task_idStringThe unique identifier for the task.e36c2221-b7fd-xxxx
output.scheduled_timeStringThe task scheduled time, in YYYY-MM-DD HH:mm:ss.SSS format.2026-01-22 10:54:44.251
output.resultsArray<Object>An array of generation results, containing metadata for the generated assets.-
output.results[].orig_promptStringThe original prompt entered by the user.a cute cat sitting on a windowsill
output.results[].actual_promptStringThe prompt used for generation after model expansion.a cute white cat with fluffy, soft fur...
output.results[].urlStringThe URL of the generated file.https://dashscope-result...
usageObjectModel usage statistics.-
usage.image_countNumberThe number of generated images (image generation mode).1
urlsArray<String>An array of URLs for the generated files. Use these URLs to access or download the output.["https://dashscope-result..."]

Test application

After you configure the workflow, you can use the test feature to verify that it runs as expected. Click the Test button in the upper-right corner to open the test panel. The test panel offers multiple test modes for different use cases.

Text conversation

Text conversation is the default test mode. It preserves the conversation history and supports continuous, multi-turn conversations.

  1. From the drop-down list at the top of the test panel, select text conversation mode (selected by default). If the workflow contains custom variables, enter their values in the parameter configuration area.

  2. In the input box, enter your test content (text and file attachments are supported), and then click the Send button or press Enter to run the test.

  3. Review the test results. You can click a node to view its detailed input and output, or switch the output format between Text and JSON.

  4. To continue the conversation, enter your next turn in the input box and send it. To start a new conversation, click the Clear All button.

Text generation

The text generation mode is for single-turn interactions. Each test is independent, and the conversation history is not preserved.

Checklist

The checklist lists the required configurations for your workflow.

To view the checklist, click the icon in the upper-right corner of the canvas configuration page.

Release an application

Once an application is released, you can call it via an API or share it as a web page with RAM users in the same main account. To do so, click the Publish button in the upper-right corner of the agent application management page.

Call via API

On the Publish Channel tab of your workflow application, click View API next to API to learn how to call the agent application by using an API.

Note: You must replace YOUR_API_KEY with your API key to call the API.

For FAQs and information about API calls, see the following topics:

  • For invocation methods (HTTP/SDK), see Application call.

  • For details on API call parameters, see Application call parameter information.

  • For details on passing parameters, see Parameter passing for applications.

  • To resolve errors returned by API calls, see Error codes.

  • The application itself has no concurrency limit; instead, the limit is determined by the models it calls. See the Model Studio console.

Currently, you cannot call the Xiyan service from a workflow. Instead, use an API node to call a custom API service.

Note

The timeout for API calls is 600 seconds and cannot be changed. If a timeout may occur, consider the following solutions:

  • Use asynchronous mode: In this mode, the system returns a task ID. You can then use the task ID to query the result, which avoids the synchronous timeout limit.

  • Split the task: Break down the task into multiple steps, or process batch data in smaller chunks to prevent a single execution from timing out.

Import or export a workflow

  1. Import or export a Model Studio workflow

Click the icon at the top of the workflow page and select Export DSL or Import Model Studio DSL.

  1. Import a Dify workflow

Model Studio supports one-click import of Dify workflows for easy migration and reuse.

  1. Click the icon at the top of the workflow page and select Import Dify DSL.

  2. Adjust the parameters for each node.

The node compatibility details are as follows:

Dify nodeMapped Model Studio nodeCompatibility
Dify nodeMapped Model Studio nodeCompatibility
startstart1. sys.query maps to query. 2. sys.dialogue_count maps to maximum memory turns.
LLMLLM1. model: Model Studio clears the model field for unsupported models, requiring you to select one manually. Supported models are fully compatible. 2. prompt: Dify's System prompt maps to the main prompt in Model Studio, while its User prompt maps to the user prompt. 3. vision capability: Fully compatible. 4. context: Model Studio incorporates the raw fields from Dify's context directly into its system prompt.
knowledge retrievalknowledge base1. input: Model Studio consistently uses the content field as the input. 2. knowledge base: Model Studio clears this field after import. You must manually associate a knowledge base in Model Studio. 3. retrieval settings: Dify's Top-k parameter maps to number of retrieved fragments.
Direct Replyoutput nodeFully compatible.
agentNoneOnly the name is retained. You must click the node and select a specific Model Studio node as a replacement.
Question Classifierintent classificationModel Studio clears the model field for unsupported models, requiring you to select one manually. Supported models are fully compatible.
Iterationbatch processing1. input: Maps to Batch Array. 2. output variable: Maps to Output Variable.
looploopFully compatible.
code executionscriptModel Studio distinguishes between Python and JavaScript scripts.
Template TransformNoneNot compatible. Model Studio generates a custom node.
Variable AggregatorVariable ProcessingMaps to the Aggregate Groups output mode of the Variable Processing node.
Document ExtractorNoneNot compatible. Model Studio generates a custom node.
Variable AssignmentVariable SettingsFully compatible.
Parameter Extractorparameter extractionModel Studio does not support inference mode. Other features are fully compatible.
HTTP requestAPIFully compatible, but you must re-authenticate.
List OperationNoneNot compatible. Model Studio generates a custom node.
Toolplugin, MCPNot compatible. Model Studio generates a custom node.
commentNoneNot compatible.
endendIf a Dify workflow contains multiple end nodes, Model Studio converts them into a single Variable Processing node and a single end node.

Manage workflow versions

1. Click Publish in the top-right corner of the workflow configuration page. In the Publish dialog box, enter version information, such as 1.0.0, and click OK.
2. Click Version Management at the top of the page, and in the Historical Versions panel, you can view or use different versions of the current workflow application as needed by clicking Overwrite Current Draft or Return to Current Version. > You can also click Export DSL for This Version at the top to export the DSL of the selected historical version.
3. Optional: In the Node Library, view or search for nodes.

Delete and copy workflow applications

In My Applications, find a published application card and click the icon to delete the application, copy its workflow, or modify its application name.

FAQ

Workflow applications

Nodes

  1. How do I write the results of a workflow run to a database?

Use a script conversion node to write the output from the previous node to a database.

  1. How do I upload files when building a workflow application in Model Studio?

Add an API node to your workflow application to upload files.

  1. How do I upload images?

Use a VL model and pass the image URL as a parameter.

  1. Can I use an asynchronous task API within a workflow application?

The timeout for a workflow application is 600 seconds. Avoid using an asynchronous task API within a workflow.

  1. How can I call the Model Studio workflow API from a frontend application and receive a streaming output?

Frontend calls are not currently supported.

  1. Why can't I import a standalone .yaml file into a Model Studio workflow?

Model Studio does not support importing standalone .yaml files. You must provide a compressed package that includes an MD5 file. We recommend regenerating the MD5 if you encounter issues.

  1. Can variable names in Model Studio workflows be in Chinese?

No, variable names cannot contain Chinese characters.

  1. How is conversation history stored?

Workflow applications store data for only one month. You must save your own conversation history. The session_id is valid for one hour.

  1. Why does my intent classification node fail when context is enabled?

If you enable context for an intent classification node, the variable that you pass to it must be a list.

  1. Why does my API node fail when I use streaming output?

The API node within a workflow does not support streaming output. However, streaming output is supported by the underlying HTTP API.

  1. How do I speed up a slow conditional node?

  2. Check your workflow configuration: Ensure each node, especially the conditional node, is configured correctly. Avoid unnecessarily complex calculations or data processing to reduce response time.

  3. Optimize script logic: If the conditional logic involves a custom script, optimize the script by reducing unnecessary loops or redundant data processing to improve performance.

  4. Run batch tests: Measure the average response time of your workflow to identify performance bottlenecks under specific conditions.

  5. How do I stream the reasoning process from a large model node?

You need to add a text conversion node after the large model node, configure the reasoning_content variable, and enable Result Return. The end node must receive the result.

  1. How can I customize the output parameters of a large model node?

  2. Use a script node to process the output: Add a script node after the large model node to process its output, convert it to your desired format, or add extra parameters.

  3. Configure a batch node: If you are using a large model node within a batch node, you can select the output of the large model node as the final output in the batch node's configuration. The steps are as follows:

    • Add a large model node to the batch node.

    • In the batch node configuration, select the output of the large model node as the final output resultList.

See Parameter passing for applications.

  1. Why does my API node return no result or fail to pass parameters?

Verify that your API key and Base URL are correct. Ensure that the input parameters are configured correctly, and adjust the field input types if necessary. Use Model Observation to view model usage details.

  1. How do I resolve issues when accessing Excel data from a knowledge base?

You cannot directly access local files. Use MCP to access local resources. You must manually process the output from a knowledge base node. We recommend adding a large model node to convert the output to a table format before passing it to a script node for further processing.

Previous: Agent applicationsNext: Prompt template overview

Is this page helpful?

Application

Why use a workflow application

Examples

Example 1: Detect scam messages

Example 2: Smart shopping assistant

Session parameters

Node

Test application

Text conversation

Text generation

Checklist

Release an application

Call via API

Import or export a workflow

Manage workflow versions

Delete and copy workflow applications

FAQ

Contact Us

Sales Support

Live-chat with our sales team or get in touch with a business development professional in your region.

Contact Sales

Technical Support

Open a ticket and get quick help from our technical team.

Open a Ticket >

Connect & Report Abuse

We look forward to your suggestion.

Post a Suggestion > Report Abuse >

\ \ Hi, I'm Alibaba Cloud AI Assistant!\ \ I can help with questions and solutions.

Why Alibaba Cloud

About Alibaba Cloud

Asia Accelerator

Our Global Network

Global Offices

Trust Center

Case Studies

Analyst Reports

Products & Pricings

Pricing Calculator

ECS

SAS

Model Studio

Database

Security

SMS

Solutions

Financial Services

Retail Services

Media Services

Gaming Services

ISV Solutions

Engage

Developer Community

Partner Network

Startups

Marketplace

Join Alibaba Cloud

Resources & Support

Developer Learning Hub

Documentation Center

Training & Certification

Service Notices

Submit a Ticket

Security Report

Qwen Cloud

Careers About Us Privacy Policy Legal Integrity Compliance Reporting Channel Service Notices Links

© 2009-2026 Copyright by Alibaba Cloud All rights reserved

© 2009-2026 Copyright by Alibaba Cloud All rights reserved

Careers About Us Privacy Policy Legal Integrity Compliance Reporting Channel Service Notices Links