Alibaba released Qwen3.8 Flash, a new multimodal 125B-parameter MoE model. > 262K native context, extensible to 1M wit

Alibaba released Qwen3.8 Flash, a new multimodal 125B-parameter MoE model. 

> 262K native context, extensible to 1M wit

Alibaba released Qwen3.8 Flash, a new multimodal 125B-parameter MoE model.

Alibaba released Qwen3.8 Flash, a new multimodal 125B-parameter MoE model. > 262K native context, extensible to 1M with YaRN. > Pricing on QwenCloud API - $ 0.16/1M input tokens and $ 0.47/1M output tokens. > Built on top of the new architecture, serving as a precursor to the architecture used in Qwen4. > Qwen3.8 Flash scores 58.7 on DeepSWE 1.1 and 62.5 on SWE-bench Pro.

Voir la source originale

À la une

ANTHROPIC 🔥: Claude limits will be permanently increased by 25% starting from September 14. > Applies for Pro, Max, T

post

threads

·

273 likes

Enter Pro

listing

·

#1 · 199 pts

This Tiny Robot Duck Learned to Walk With Reinforcement Learning Microduck is the tiny biped duck robot that actually

video

instagram

robotics

reinforcementlearning

embodiedai

·

112,7 k vues

PageIndex

listing

·

#1 · 177 pts

Dans la même veine

ANTHROPIC 🔥: Claude limits will be permanently increased by 25% starting from September 14. 

> Applies for Pro, Max, T

post

threads

·

273 likes

ANTHROPIC 🔥: Claude limits will be permanently increased by 25% starting from September 14. > Applies for Pro, Max, T

post

threads

·

168 likes

OPENAI 👀: ChatGPT got a bunch of updates recently. Users can now automatically generate q…

ZHIPU AI 🔥: GLM-5.3-Flash (Ox Alpha) has been officially announced!

> GLM-5.3-Flash is a 320B-A18B multimodal model.

post

threads

·

89 likes

ZHIPU AI 🔥: GLM-5.3-Flash (Ox Alpha) has been officially announced! > GLM-5.3-Flash is a 320B-A18B multimodal model.