Qwen 3.8 Flash
Model Information
Display Name: Qwen 3.8 Flash
API Model ID: qwen/qwen3.8-flash
Category: Image To Text
Description: Qwen 3.8 Flash is Alibaba's fast multimodal reasoning model — the Qwen 3.8 generation of the Flash line. It reads text, images and video, switches reasoning on and off per request, and is tuned for agents, visual coding, document and long-video analysis at a fraction of the flagship cost. **Key Features:** - 1M token context window - Up to 131K output tokens - Vision + video input (text output) - Hybrid thinking: off by default, on via parameter (up to 262K reasoning tokens) - Function/tool calling and structured outputs (JSON mode) - Prompt caching for cheaper repeated context - Streaming support **Best For:** - High-volume multimodal workloads - Visual agents, UI automation and document understanding - Long-context and long-video processing on a tight budget - Fast classification, extraction and routing
Context Window: 1,000,000 tokens
Max Output: 131,072 tokens
How to Use This Model
To use Qwen 3.8 Flash via the HInow.ai API, use the model ID: qwen/qwen3.8-flash
API Request Example (Chat/Text)
POST https://api.hinow.ai/v1/chat/completions
Authorization: Bearer YOUR_API_KEY
Content-Type: application/json
{
"model": "qwen/qwen3.8-flash",
"messages": [
{"role": "user", "content": "Your message here"}
]
}
API Request Example (Image Generation)
POST https://api.hinow.ai/v1/images
Authorization: Bearer YOUR_API_KEY
Content-Type: application/json
{
"model": "qwen/qwen3.8-flash",
"prompt": "Your image description here"
}
Pricing
- input: $1.50
- output: $5.00
- cached: $0.15
Available Parameters
- temperature: Controls randomness (0-2). Default: 0.7 (Options: 0, 0.3, 0.5, 0.7, 1.0, 1.5, 2.0)
- top_p: Nucleus sampling (0-1). Default: 0.9 (Options: 0.1, 0.5, 0.7, 0.9, 0.95, 1.0)
- max_tokens: Max tokens to generate (1-131072) (Options: 256, 512, 1024, 2048, 4096, 8192, 16384, 32768, 65536, 131072)
- thinking: Chain-of-thought reasoning. Default: off (Options: on, off)
- response_format: Output format (Options: text, json_object)
- seed: Deterministic sampling seed
- presence_penalty: Penalizes repeated topics (-2 to 2). Default: 0 (Options: -2.0, -1.0, 0, 1.0, 2.0)
- tools: Function/tool definitions for agentic workflows
Quick Reference
To use this model, set: "model": "qwen/qwen3.8-flash"
Featured: No
Documentation: https://hinow.ai/models/qwen/qwen3.8-flash
API Endpoint: https://api.hinow.ai/v1


