nexusflow
Online

Models

59 models

Browse the full range of AI models, covering text, reasoning, vision, coding, image, video, embedding and more

Qwen3.7 PlusNEWHOT
Qwen (Alibaba)Multimodal Model
qwen3.7-plus
OpenAIAnthropicResponses

Qwen3.7 series cost-effective Plus model, with fully upgraded visual-language capabilities on top of strong text abilities, retaining complete agent capabilities for coding, tool use, and productivity workflows. Supports multimodal interactive hybrid agents: perceiving real-world scenes, reading screens and operating GUIs, generating code based on visual references, and end-to-end navigation of mobile apps. Equivalent to snapshot qwen3.7-plus-2026-05-26.

cost-effectivemultimodalagentvision
Context
1M
Input (tier 1)
$0.29/M
Output (tier 1)
$1.14/M
Qwen3.7 MaxNEWHOT
Qwen (Alibaba)Language Model
qwen3.7-max

Qwen3.7 flagship model, built for the agent era with comprehensive improvements in coding, office productivity, and long-cycle autonomous execution. Supports thinking mode toggle, function calling, and web search. 1M token context window.

flagshipreasoningcodingthinking mode
Context
1M
Input
$1.71/M
Output
$5.14/M
Qwen3 MaxNEWHOT
Qwen (Alibaba)Language Model
qwen3-max

Qwen3 most powerful flagship model, supports thinking mode toggle, excels in complex reasoning, code generation, and mathematics. 262K context window.

flagshipreasoningcodingthinking mode
Context
256K
Input (tier 1)
$0.36/M
Output (tier 1)
$1.43/M
Qwen3.6 Max PreviewNEWHOT
Qwen (Alibaba)Language Model
qwen3.6-max-preview

Qwen3.6 most powerful preview model, designed for complex reasoning, code generation, and multi-step tool tasks, ideal for scenarios requiring stronger thinking capabilities.

flagshipreasoningpreview
Context
256K
Input (tier 1)
$1.29/M
Output (tier 1)
$7.71/M
Qwen3.6 PlusNEWHOT
Qwen (Alibaba)Language Model
qwen3.6-plus

Qwen3.6 balanced flagship model, supports 1M context window, function calling, and built-in tools, ideal for large codebases and general production scenarios.

cost-effectivebalanced1M context
Context
1M
Input (tier 1)
$0.29/M
Output (tier 1)
$1.71/M
Qwen3.5 PlusNEWHOT
Qwen (Alibaba)Language Model
qwen3.5-plus

Qwen3.5 enhanced model, best balance of quality, speed, and cost. Supports 1M context window, ideal for large-scale application scenarios.

cost-effectivebalanced1M context
Context
1M
Input (tier 1)
$0.11/M
Output (tier 1)
$0.69/M
Qwen3.6 FlashNEW
Qwen (Alibaba)Language Model
qwen3.6-flash

Qwen3.6 flash model, ideal for simple tasks with fast speed and low cost. Supports 1M context window and context caching.

ultra-fastlow cost1M context
Context
1M
Input (tier 1)
$0.17/M
Output (tier 1)
$1.03/M
Qwen3.5 FlashNEW
Qwen (Alibaba)Language Model
qwen3.5-flash

Qwen3.5 flash model, ideal for simple tasks with fast speed and low cost. Supports 1M context window and context caching.

ultra-fastlow cost1M context
Context
1M
Input (tier 1)
$0.03/M
Output (tier 1)
$0.29/M
Qwen PlusHOT
Qwen (Alibaba)Language Model
qwen-plus

Qwen enhanced model, classic balance of quality and speed, ideal for large-scale application scenarios.

cost-effectivebalancedgeneral
Context
1M
Input (tier 1)
$0.11/M
Output (tier 1)
$0.29/M
Qwen Turbo
Qwen (Alibaba)Language Model
qwen-turbo

Qwen high-speed model, extremely fast response and lowest cost, ideal for latency-sensitive application scenarios.

fastlow costgeneral
Context
1M
Input
$0.04/M
Output
$0.09/M
Qwen Long
Qwen (Alibaba)Language Model
qwen-long

Qwen long-context model, supports ultra-long context windows up to 10M tokens, ideal for document analysis and long-text understanding.

ultra-long contextdocument analysis
Context
10M
Input
$0.07/M
Output
$0.28/M
Qwen FlashNEW
Qwen (Alibaba)Language Model
qwen-flash

Qwen ultra-fast general model, 1M context window, extremely fast response and ultra-low cost, ideal for large-scale high-concurrency scenarios. Supports function calling and thinking mode.

ultra-fastlow cost1M contextgeneral
Context
1M
Input
$0.02/M
Output
$0.21/M
Qwen3 235B-A22BNEW
Qwen (Alibaba)Language Model
qwen3-235b-a22b

Qwen3 open-source flagship, 235B parameter MoE architecture (22B active), supports dynamic switching between thinking and non-thinking modes.

open-sourceMoEreasoningthinking mode
Context
128K
Input
$0.29/M
Output
$1.14/M
Qwen3.6 35B-A3BNEW
Qwen (Alibaba)Language Model
qwen3.6-35b-a3b

Qwen3.6 open-source MoE model, 35B total parameters with only 3B active, excels at agent coding, STEM, and reasoning tasks. Apache 2.0 licensed. Supports thinking mode toggle.

open-sourceMoElightweightcoding
Context
256K
Input
$0.26/M
Output
$1.54/M
Qwen3 32B
Qwen (Alibaba)Language Model
qwen3-32b

Qwen3 open-source 32B parameter dense model, excels among medium-scale models.

open-sourcereasoningcoding
Context
128K
Input
$0.29/M
Output
$1.14/M
QwQ PlusHOT
Qwen (Alibaba)Reasoning Model
qwq-plus
OpenAIAnthropicResponses

Qwen reasoning model, trained on Qwen2.5, excels at mathematics, logical reasoning, and complex problem analysis, displaying complete chain-of-thought.

reasoningmathematicslogicchain-of-thought
Context
128K
Input
$0.23/M
Output
$0.57/M
Qwen VL MaxHOT
Qwen (Alibaba)Multimodal Model
qwen-vl-max
OpenAIAnthropicResponses

Qwen vision flagship model, supports image understanding, visual-text dialogue, document OCR, and other multimodal tasks.

visionmultimodalOCRvisual understanding
Context
128K
Input
$0.23/M
Output
$0.57/M
Qwen VL Plus
Qwen (Alibaba)Multimodal Model
qwen-vl-plus
OpenAIAnthropicResponses

Qwen vision enhanced model, balanced performance and cost multimodal model.

visionmultimodalcost-effective
Context
128K
Input
$0.11/M
Output
$0.29/M
Qwen3 VL PlusNEW
Qwen (Alibaba)Multimodal Model
qwen3-vl-plus
OpenAIAnthropicResponses

Qwen3 vision-language model, significantly improved image understanding, supports high-resolution image input. 262K context window.

visionmultimodalhigh-resolution
Context
256K
Input (tier 1)
$0.14/M
Output (tier 1)
$1.43/M
Qwen3 VL Flash
Qwen (Alibaba)Multimodal Model
qwen3-vl-flash
OpenAIAnthropicResponses

Qwen3 vision flash model, fast image understanding, ideal for real-time scenarios.

visionultra-fastcost-effective
Context
256K
Input (tier 1)
$0.02/M
Output (tier 1)
$0.21/M
Qwen3.5 Omni PlusNEWHOT
Qwen (Alibaba)Multimodal Model
qwen3.5-omni-plus
OpenAIAnthropicResponses

Qwen3.5 flagship omni model, supports any combination of text, image, audio, and video input with text and voice output. Up to 3 hours audio / 1 hour video input, 113 input languages, 55 voice tones, supports web search and voice cloning.

flagshipomnimultimodalaudio input
Context
256K
Input
$0.97/M
Output
$5.52/M
Qwen3.5 Omni FlashNEWHOT
Qwen (Alibaba)Multimodal Model
qwen3.5-omni-flash
OpenAIAnthropicResponses

Qwen3.5 lightweight omni model, supports any combination of text, image, audio, and video input with text and voice output. Up to 3 hours audio / 1 hour video input, 113 input languages, 55 voice tones, supports web search. Best cost-efficiency choice.

cost-effectiveomnimultimodalaudio input
Context
256K
Input
$0.3/M
Output
$1.83/M
Qwen3 Omni Flash
Qwen (Alibaba)Multimodal Model
qwen3-omni-flash
OpenAIAnthropicResponses

Qwen3 omni model, accepts text, image, audio, and video inputs with text and voice output. Supports thinking mode (text-only output in thinking mode). Ideal for short video analysis and cost-sensitive scenarios.

omnimultimodalaudio inputaudio output
Context
64K
Input
$0.26/M
Output
$0.99/M
Qwen3 Coder PlusNEWHOT
Qwen (Alibaba)Code Model
qwen3-coder-plus

Qwen3 exceptional code model, excels at tool calling and environment interaction, with outstanding code generation, completion, debugging, and refactoring capabilities. 1M context window.

codingcode generationtool calling1M context
Context
1M
Input (tier 1)
$0.57/M
Output (tier 1)
$2.29/M
Qwen3 Coder FlashNEW
Qwen (Alibaba)Code Model
qwen3-coder-flash

Qwen3 code flash model, fast code completion and generation, ideal for IDE integration scenarios.

codingultra-fastcost-effective
Context
1M
Input (tier 1)
$0.14/M
Output (tier 1)
$0.57/M
Qwen Math PlusNEW
Qwen (Alibaba)Reasoning Model
qwen-math-plus
OpenAIAnthropicResponses

Qwen math-specialized model, excels at mathematical problem solving, proofs, and computation, supports LaTeX format output.

mathematicsreasoningproblem solvingLaTeX
Context
4K
Input
$0.55/M
Output
$1.66/M
Qwen MT PlusNEW
Qwen (Alibaba)Specialty Model
qwen-mt-plus

Qwen flagship translation model, supports 92 language pairs, outstanding translation quality, ideal for professional translation scenarios.

translation92 languagesprofessional
Context
16K
Input
$0.25/M
Output
$0.74/M
Tongyi Intent Detect V3NEW
Qwen (Alibaba)Specialty Model
tongyi-intent-detect-v3

Qwen intent understanding model, rapidly and accurately parses user intent within milliseconds, suitable for customer service routing, intelligent dialogue distribution, and instruction parsing scenarios.

intent detectionfastcustomer service routingclassification
Context
8K
Input
$0.06/M
Output
$0.14/M
Text Embedding V4NEWHOT
Qwen (Alibaba)Embedding Model
text-embedding-v4
Embedding

Qwen latest text embedding model, supports 100+ languages and multiple programming languages, vector dimensions selectable from 2048, 1536, 1024, 768, 512, 256, 128, 64, suitable for semantic retrieval, clustering, recommendation, and RAG.

embeddingvectorsemantic searchRAG
Context
8K
Input
$0.07/M
Output
Free
Text Embedding V3NEW
Qwen (Alibaba)Embedding Model
text-embedding-v3
Embedding

Qwen text embedding model, converts text into high-dimensional vector representations, suitable for semantic search, clustering, and recommendation scenarios.

embeddingvectorsemantic search
Context
8K
Input
$0.07/M
Output
Free
Qwen3 ASR FlashNEW
Qwen (Alibaba)Audio Model
qwen3-asr-flash
ASR

Qwen3 speech recognition model, supports automatic detection and transcription in 11 languages, word-level timestamps, emotion recognition, singing recognition, and speaker diarization. Supports both real-time and non-real-time modes.

speech recognitionASRmultilingualreal-time
Context
0
Price
$0.03/s
Billing
Per second
Qwen3 TTS Flash RealtimeNEW
Qwen (Alibaba)Audio Model
qwen3-tts-flash-realtime
TTS

Qwen3 real-time text-to-speech model, streaming synthesis via WebSocket, supports multiple languages and voice tones including Chinese and English, suitable for voice assistants and audiobooks.

text-to-speechTTSreal-timemultilingual
Context
0
Price
$0.14/s
Billing
Per second
Wan 2.6 Text-to-ImageNEWHOT
Qwen (Alibaba)Image Generation
wan2.6-t2i
ImageTasks

Latest generation text-to-image flagship model, supports mixed text-image output and image editing. Can process complex instructions, render Chinese and English text, and generate high-definition realistic images. Supports multiple resolutions and aspect ratios.

image generationtext-to-imagemixed text-imageHD realistic
Context
4K
Price
$0.03/image
Billing
Per image
Wan 2.6 Text-to-VideoNEWHOT
Qwen (Alibaba)Video Generation
wan2.6-t2v
Tasks

Latest generation text-to-video flagship model, supports multi-shot narrative and intelligent storyboard. Can generate 2-15 second 1080P HD video, supports prompt rewriting. Generation time approximately 1-5 minutes.

video generationtext-to-videomulti-shot1080P
Context
1K
Price
$0.09/s
Billing
Per second
Wan 2.6 Image-to-VideoNEWHOT
Qwen (Alibaba)Video Generation
wan2.6-i2v
Tasks

Image-driven video generation model, uses the input image as the first frame to generate coherent video. Supports multi-shot narrative, automatic dubbing, 720P/1080P resolution, 2-15 seconds duration. Excellent frame coherence and motion consistency.

video generationimage-to-videofirst-frame drivenmulti-shot
Context
1K
Price
$0.09/s
Billing
Per second
Wan 2.6 Image-to-Video FlashNEW
Qwen (Alibaba)Video Generation
wan2.6-i2v-flash
Tasks

Fast image-to-video model, supports audio and silent video generation. Faster generation speed, ideal for latency-sensitive scenarios. Supports 720P/1080P, 2-15 seconds duration.

video generationimage-to-videofastFlash
Context
1K
Price
$0.02/s
Billing
Per second
Wan 2.6 Reference-to-VideoNEW
Qwen (Alibaba)Video Generation
wan2.6-r2v
Tasks

Multimodal input video generation model, supports text/image/video as references. Can use characters or objects as protagonists to generate single-character or multi-character interaction videos. 2-10 seconds duration, supports intelligent storyboarding.

video generationreference-to-videorole playmultimodal
Context
1K
Price
$0.08/s
Billing
Per second
Wan 2.6 Reference-to-Video FlashNEW
Qwen (Alibaba)Video Generation
wan2.6-r2v-flash
Tasks

Fast reference-to-video model, supports audio and silent output. Faster generation speed, ideal for rapid iteration scenarios. Supports 720P/1080P resolution.

video generationreference-to-videofastFlash
Context
1K
Price
$0.02/s
Billing
Per second
PixVerse V6NEWHOT
PixVerseVideo Generation
pixverse-v6
Tasks

PixVerse latest flagship video generation model, supports text-to-video and image-to-video, with significantly improved visual quality and motion consistency. Supports 1-15 seconds duration, 360p/540p/720p/1080p multiple resolutions, various aspect ratios.

video generationtext-to-videoimage-to-videoflagship
Context
500
Price
$0.02/s
Billing
Per second
HappyHorse 1.0 Text-to-VideoNEWHOT
AlibabaVideo Generation
happyhorse-1.0-t2v
Tasks

Alibaba's 2026 latest AI video generation model, ranked #1 on benchmarks. Generates high-quality video from text, supports 720P/1080P, 3-15 seconds duration, various aspect ratios. Default audio included.

video generationtext-to-videohigh qualitybenchmark #1
Context
2K
Price
$0.13/s
Billing
Per second
HappyHorse 1.0 Image-to-VideoNEWHOT
AlibabaVideo Generation
happyhorse-1.0-i2v
Tasks

Generates coherent video using the input image as the first frame, supports 720P/1080P, 3-15 seconds duration. Excellent frame coherence and motion consistency. Default audio included.

video generationimage-to-videofirst-frame drivenhigh quality
Context
2K
Price
$0.13/s
Billing
Per second
HappyHorse 1.0 Reference-to-VideoNEW
AlibabaVideo Generation
happyhorse-1.0-r2v
Tasks

Supports 1-9 reference images input, can fuse characters/objects/scenes from images to generate video. Supports 720P/1080P, 3-15 seconds, various aspect ratios. Default audio included.

video generationreference-to-videomulti-image inputhigh quality
Context
2K
Price
$0.13/s
Billing
Per second
HappyHorse 1.0 Video EditNEW
AlibabaVideo Generation
happyhorse-1.0-video-edit
Tasks

AI video editing based on input video, supports 0-5 reference images for assisted editing. Input video 3-60 seconds (truncated beyond 15s), supports 720P/1080P, can preserve original audio.

video generationvideo editingAI editingaudio preservation
Context
2K
Price
$0.13/s
Billing
Per second
DeepSeek V4 FlashNEW
DeepSeekLanguage Model
deepseek-v4-flash

DeepSeek V4 Flash high-speed model via DashScope, ideal for low-latency and high-concurrency online dialogue scenarios.

V4ultra-fasthigh concurrencycost-effective
Context
1M
Input
$0.14/M
Output
$0.29/M
DeepSeek V4 ProNEWHOT
DeepSeekReasoning Model
deepseek-v4-pro
OpenAIAnthropicResponses

DeepSeek V4 Pro flagship model via DashScope, designed for complex reasoning, code generation, and multi-step tasks.

V4flagshipreasoningcoding
Context
1M
Input
$1.71/M
Output
$3.43/M
DeepSeek V3.2NEWHOT
DeepSeekLanguage Model
deepseek-v3.2

DeepSeek latest general-purpose LLM, MoE architecture, strong bilingual Chinese-English capabilities, powerful coding abilities.

MoEcodingmultilingual
Context
128K
Input
$0.29/M
Output
$0.43/M
DeepSeek R1HOT
DeepSeekReasoning Model
deepseek-r1
OpenAIAnthropicResponses

DeepSeek reasoning model with outstanding performance in mathematics, coding, and logical reasoning, displaying complete chain-of-thought.

reasoningmathematicscodingchain-of-thought
Context
128K
Input
$0.55/M
Output
$2.21/M
DeepSeek V3
DeepSeekLanguage Model
deepseek-v3

DeepSeek V3 general-purpose LLM, 671B parameter MoE architecture, excellent bilingual Chinese-English capabilities.

MoEmultilingualcoding
Context
128K
Input
$0.28/M
Output
$1.1/M
Claude Opus 4.7NEWHOT
AnthropicLanguage Model
claude-opus-4-7
Anthropic

Anthropic's most capable general-purpose model, designed for complex reasoning, agentic coding, and long-context tasks. Official pricing: $5 input / $25 output per 1M tokens.

Claudeflagshipagentvision
Context
1M
Input
$5/M
Output
$25/M
Claude Sonnet 4.6NEWHOT
AnthropicLanguage Model
claude-sonnet-4-6
Anthropic

Anthropic's balanced speed and intelligence model, ideal for production-grade dialogue, code, tool use, and long-context workflows. Official pricing: $3 input / $15 output per 1M tokens.

Claudebalancedcodingvision
Context
1M
Input
$3/M
Output
$15/M
Claude Haiku 4.5NEW
AnthropicLanguage Model
claude-haiku-4-5
Anthropic

Anthropic's fast and low-cost model with near-frontier intelligence, ideal for low-latency dialogue, classification, extraction, and batch tasks. Official pricing: $1 input / $5 output per 1M tokens.

Claudeultra-fastlow costvision
Context
195K
Input
$1/M
Output
$5/M
GLM 4.7NEW
Zhipu AILanguage Model
glm-4.7

Zhipu AI latest GLM-4.7 model with significantly improved overall capabilities and strong Chinese language understanding.

Chinese optimizedreasoninggeneral
Context
166K
Input (tier 1)
$0.41/M
Output (tier 1)
$1.93/M
GLM 5NEWHOT
Zhipu AILanguage Model
glm-5

Zhipu AI GLM-5 flagship model with comprehensive capability improvements, outstanding performance in reasoning, coding, and long-text tasks.

flagshipreasoningcodingChinese optimized
Context
198K
Input (tier 1)
$0.55/M
Output (tier 1)
$2.48/M
GLM 5.1NEWHOT
Zhipu AILanguage Model
glm-5.1

Zhipu AI GLM-5.1 enhanced flagship model, further optimized over GLM-5, stronger complex reasoning and code generation capabilities.

flagshipreasoningcodingenhanced
Context
198K
Input (tier 1)
$0.83/M
Output (tier 1)
$3.31/M
Kimi K2.5
Moonshot AILanguage Model
kimi-k2.5

Moonshot AI Kimi K2.5 model, excels at long-text understanding and multi-turn dialogue with outstanding Chinese language capabilities.

long contextmulti-turn dialogueChinese optimized
Context
256K
Input
$0.55/M
Output
$2.9/M
Kimi K2.6NEWHOT
Moonshot AILanguage Model
kimi-k2.6

Moonshot AI Kimi K2.6 latest flagship model with significantly improved long-text understanding and creative writing, supports longer context windows.

flagshiplong contextcreative writingChinese optimized
Context
256K
Input
$0.9/M
Output
$3.72/M
MiniMax M2.1
MiniMaxLanguage Model
MiniMax-M2.1

MiniMax M2.1 model, outstanding performance in creative writing and multi-turn dialogue.

creative writingdialoguegeneral
Context
200K
Input
$0.29/M
Output
$1.16/M
MiniMax M2.5NEW
MiniMaxLanguage Model
MiniMax-M2.5

MiniMax M2.5 enhanced model with improved reasoning and coding capabilities, more stable multi-turn dialogue.

reasoningcodingdialogue
Context
192K
Input
$0.29/M
Output
$1.16/M
Qwen3 8BNEW
Qwen (Alibaba)Language Model
qwen3-8b

Qwen3 open-source 8B parameter lightweight model, ideal for edge deployment and low-cost inference scenarios.

open-sourcelightweightcost-effective
Context
128K
Input
$0.07/M
Output
$0.29/M