Alibaba
+1 sub-topics in scope Saves to local browser storage. Followed topics appear on the homepage and refresh on each visit.Also known as alibaba group·alibaba cloud·alicloud·qwen·qwen 2.5
Latest from across the web
External coverage we have crawled and indexed for this topic.
I gave Qwen 3.8 27B a reverse-engineering job I assumed needed a frontier model, and it finished in 30 minutes
Qwen 3.8 27B genuinely shocked me with what it achieved here.
Rent a supernode by the hour - Alibaba brings frontier-scale AI compute to the public cloud, but only if you live in this remote Chinese province
A homegrown AI supernode for hire, for domestic users only, for now
Alibaba's TSMC-Built 5nm RISC-V Chip, XuanTie C950, Now Runs Qwen-3.8 27B Model Natively, Unlocking Massive Vertical Integration Tailwinds
Alibaba's XuanTie C950 chip hits 30 tokens/second, with a 1.9-second first-token latency, on Qwen-3.8 27B - no emulation layer required.
DeepSeek's Peak-Hour Pricing Betrays Where Its Users Really Live, Calming US Fears of a China AI Takeover, Even As Alibaba's Qwen Models Bury Meta On Hugging Face
DeepSeek's surge windows fall between 09:00 and 16:00 Beijing time — a tell that most demand still sits inside Asian time zones.
Report: Alibaba to sell gaming division for over $2 billion | Game World Observer
It is allegedly going to be acquired by the investment firm Trustar Capital.
Lokale KI ausprobiert: Das leistet Qwen3.8-27B
Alibaba hat die Gewichte seines Modells Qwen3.8-27B veröffentlicht, klein genug für bezahlbare Grafikkarten. Der Abstand zu Spitzenmodellen schrumpft.
Adjacent signals
Latest from topics that share context with Alibaba — parents, siblings, descendants.
Related in graph
Videos
From the channels we trackDiscussions on the web
Recent threads on Reddit and Hacker News that mention Alibaba.
People also ask
Common questions on Alibaba, surfaced from across the indexed web.
What is Alibaba offering to enterprise consumers?
A supernode is a set of accelerators wired together tightly enough that software treats them as one very large chip rather than a network of separate ones. This approach is often necessary; a model too big to fit on a single card would otherwise have to shovel data across comparatively slow networking, and that overhead can determine whether trillion-parameter models are practical to serve. One of the best examples is Nvidia's GB200 NVL72, which leverages 72 of its Blackwell GPUs in a single configuration, or AMD's 72-GPU Helios configuration, which offers its own Instinct MI455X GPUs with mor
Rent a supernode by the hour - Alibaba brings frontier-scale AI compute to the public cloud, but only if you live in this remote Chinese provinceCan you copy an AI just by talking to it?
Model distillation is a common technique used by AI companies to build variations of their models, especially smaller, faster options. But no company would be okay with a rival using their model to train the competition. But that's what Anthropic alleges. The fake accounts supposedly asked Claude a ton of very complex and detailed questions related to its advanced software engineering and agentic reasoning features. The responses filled in a picture of the model's workings, accelerating Alibaba's own development of competing AI systems, Anthropic claimed. The conundrum is obvious. Large langua
Anthropic says Alibaba may have copied Claude just by asking questions