Recent signals
-
Making Qwen 3.8 27B fast on Strix Halo gfx1151
Level1Techs Forum
The headlines don’t reference Alibaba directly; instead, they focus on Qwen3.8-Flash-Next and related 27B/Qwen customization and performance efforts across forums and community posts. Discussions center on running and optimizing Qwen 3.8 models (including GPU/compute tweaks) rather than any Alibaba-specific product news.
Also known as alibaba group·alibaba cloud·alicloud·qwen·qwen 2.5
External coverage we have crawled and indexed for this topic.
DeepSeek's surge windows fall between 09:00 and 16:00 Beijing time — a tell that most demand still sits inside Asian time zones.
Qwen 3.8 27B genuinely shocked me with what it achieved here.
Alibaba's XuanTie C950 chip hits 30 tokens/second, with a 1.9-second first-token latency, on Qwen-3.8 27B - no emulation layer required.
Alibaba hat die Gewichte seines Modells Qwen3.8-27B veröffentlicht, klein genug für bezahlbare Grafikkarten. Der Abstand zu Spitzenmodellen schrumpft.
A homegrown AI supernode for hire, for domestic users only, for now
It is allegedly going to be acquired by the investment firm Trustar Capital.
Source-by-source view of what publications and communities are surfacing right now.
Tracking: Make a custom qwen 3.8 27B abliterated for ninfer? / Making Qwen 3.8 27B fast on Strix Halo gfx1151
Tracking: Qwen3.8-Flash-Next
Latest from topics that share context with Alibaba — parents, siblings, descendants.
Recent threads on Reddit and Hacker News that mention Alibaba.
Common questions on Alibaba, surfaced from across the indexed web.
A supernode is a set of accelerators wired together tightly enough that software treats them as one very large chip rather than a network of separate ones. This approach is often necessary; a model too big to fit on a single card would otherwise have to shovel data across comparatively slow networking, and that overhead can determine whether trillion-parameter models are practical to serve. One of the best examples is Nvidia's GB200 NVL72, which leverages 72 of its Blackwell GPUs in a single configuration, or AMD's 72-GPU Helios configuration, which offers its own Instinct MI455X GPUs with mor
Rent a supernode by the hour - Alibaba brings frontier-scale AI compute to the public cloud, but only if you live in this remote Chinese provinceModel distillation is a common technique used by AI companies to build variations of their models, especially smaller, faster options. But no company would be okay with a rival using their model to train the competition. But that's what Anthropic alleges. The fake accounts supposedly asked Claude a ton of very complex and detailed questions related to its advanced software engineering and agentic reasoning features. The responses filled in a picture of the model's workings, accelerating Alibaba's own development of competing AI systems, Anthropic claimed. The conundrum is obvious. Large langua
Anthropic says Alibaba may have copied Claude just by asking questions