Qwen3.8-Flash-NextAlibaba
|
||||||
Related Products
|
||||||
About
Qwen3.8-Flash-Next is an open-weight multimodal Mixture-of-Experts model and an early preview of the architecture planned for Qwen4. It systematically upgrades attention, residual connections, embeddings, and optimization to improve capability, computational efficiency, model capacity, and training stability. Its hybrid architecture combines Gated DeltaNet, which efficiently compresses historical information, with Qwen Sparse Attention, which selects important context at the micro-block level to reduce attention and indexing costs on long sequences. Gated Residual widens the residual stream into four branches and dynamically controls information flow across layers, while N-gram Embedding adds large-scale local-pattern memory with very little extra per-token computation and can be offloaded to host memory. The model uses a 125B-parameter main network plus 51B N-gram embedding parameters, while activating only 6B parameters per token.
|
About
Vynaris gives teams powerful hosted uncensored models for authorized security testing, red-teaming, and research. Qwen3.8-27B, DeepSeek-V4-Flash-0731, and Qwen3.6-35B-A3B are available through an OpenAI-compatible API with published token rates and no prompt or output retention for these hosted models. Vynaris also routes requests across a broader model catalog and shows transparent per-request cost receipts, so applications can switch with a base URL change and inspect the cost of individual requests.
|
|||||
Platforms Supported
Windows
Not Supported
Mac
Not Supported
Linux
Not Supported
Cloud
Supported
On-Premises
Not Supported
iPhone
Not Supported
iPad
Not Supported
Android
Not Supported
Chromebook
Not Supported
|
Platforms Supported
Windows
Not Supported
Mac
Not Supported
Linux
Not Supported
Cloud
Supported
On-Premises
Not Supported
iPhone
Not Supported
iPad
Not Supported
Android
Not Supported
Chromebook
Not Supported
|
|||||
Audience
Developers, researchers, and AI teams seeking to run or study an efficient multimodal open-weight model with long-context reasoning, coding, multilingual, and agentic capabilities
|
Audience
Security researchers, red teams and developers who need hosted uncensored models or an OpenAI-compatible LLM gateway
|
|||||
Support
Phone Support
Not Supported
24/7 Live Support
Not Supported
Online
Supported
|
Support
Phone Support
Not Supported
24/7 Live Support
Not Supported
Online
Not Supported
|
|||||
API
Offers API
Supported
|
API
Offers API
Not Supported
|
|||||
Screenshots and Videos |
Screenshots and VideosNo images available
|
|||||
Pricing
$2 per 1M (input)
Free Version
Not Supported
Free Trial
Not Supported
|
Pricing
$5 minimum credit top-up
Free Version
Not Supported
Free Trial
Not Supported
|
|||||
Reviews/
|
Reviews/
|
|||||
Training
Documentation
Supported
Webinars
Not Supported
Live Online
Not Supported
In Person
Not Supported
|
Training
Documentation
Not Supported
Webinars
Not Supported
Live Online
Not Supported
In Person
Not Supported
|
|||||
Company InformationAlibaba
Founded: 1999
China
qwen.ai/blog
|
Company InformationVynaris
Founded: 2025
United Kingdom
vynaris.com
|
|||||
Alternatives |
Alternatives |
|||||
|
|
|
|||||
|
|
||||||
|
|
||||||
|
|
||||||
Categories |
Categories |
|||||
Integrations
Alibaba Cloud
Supported
Alibaba Cloud Model Studio
Supported
Cherry Studio
Supported
Cline
Supported
ClinePass
Supported
Happy Shrimp 1.0
Supported
Hermes Agent
Supported
Hugging Face
Supported
Model Context Protocol (MCP)
Supported
Novita AI
Supported
|
Integrations
Alibaba Cloud
Not Supported
Alibaba Cloud Model Studio
Not Supported
Cherry Studio
Not Supported
Cline
Not Supported
ClinePass
Not Supported
Happy Shrimp 1.0
Not Supported
Hermes Agent
Not Supported
Hugging Face
Not Supported
Model Context Protocol (MCP)
Not Supported
Novita AI
Not Supported
|
|||||
|
|
|