
Qwen 3.8 Flash Next GGUF
Fast long-context chat in an optimized GGUF build
Unsloth's optimized Qwen 3.8 Flash Next GGUF deployment for long-context chat, coding, research, and optional thinking-mode responses.
Live playground
Try Qwen 3.8 Flash Next GGUF in your browser
Run a real completion using your imageat account. We reserve 10 credits, charge only the completed run's actual usage tier, and automatically return the rest.
Output
Your response will appear here
Enter a prompt and run the model.
Qwen 3.8 Flash Next GGUF API pricing
Wiro's provider rate is $0.000585 per execution second. imageat applies a transparent 2× multiplier, making the API rate $0.00117 per execution second. Final API charges are rounded up to whole cents.
Provider rate
$0.000585 / sec
imageat API · 2×
$0.00117 / sec
Playground reservation
10 credits max
Reason with context
Use the model for analysis, planning, and multi-step knowledge work.
Build and code
Apply it to software development, structured output, and agent workflows.
Generate text
Create clear answers, summaries, drafts, and conversational experiences.
Model overview
What Qwen 3.8 Flash Next GGUF is built for
Unsloth's optimized Qwen 3.8 Flash Next GGUF deployment for long-context chat, coding, research, and optional thinking-mode responses.
- Long-context conversational work
- Optional visible thinking mode
- Coding, research, and document analysis
- Adjustable sampling and output controls