Dragonfly v2.5.0 is released
…Users can run commands such as dfget hf://deepseek-ai/DeepSeek-OCR and dfget modelscope://models/deepseek-ai/DeepSeek-OCR to fetch repositories. Git LFS data is downloaded through Dragonfly P2P acceleration…
Tracked topic
Another DeepSeek Moment Has Arrived
DeepSeek V4 Pro 0813 with Major Agent Upgrade: Tested Locally
DeepSeek Just Made Closed AI Look Ridiculous
DeepSeek's Absolutely Insane AI Speed Hack (DSpark)
DeepSeek Just Solved AI's Billion Dollar Problem
DeepSeek is back... and Silicon Valley is terrified
DeepSeek V4 AI Beats Billion Dollar Systems…For Free
DeepSeek Just Fixed One Of The Biggest Problems With AI
My Honest Thoughts about Deepseek
DeepSeek’s New AI Is A Game Changer
GPT 5.5 Arrives, DeepSeek V4 Drops, and the Compute War Intensifies
Deepseek ban?
…Users can run commands such as dfget hf://deepseek-ai/DeepSeek-OCR and dfget modelscope://models/deepseek-ai/DeepSeek-OCR to fetch repositories. Git LFS data is downloaded through Dragonfly P2P acceleration…
…These latest cuts bring Luna into the realm of DeepSeek V4, with its pro model costing $0.435 per million input tokens and $0.87 per million output tokens. GPT 5.6…
…GPT-5.5 led with a 70% solve rate, DeepSeek V4 Pro solved it for $0.62 per attempt, and Gemini refused to engage almost entirely.
…The range of supported models is equally wide, covering GPT, Gemini , Claude, DeepSeek, Grok, Kimi K2, and Qwen 3 on the chat side alone. The creative side is where things get particularly…
DeepSeek: What They Invented
Cosyncing: Control DeepSeek Harness and other agents on your phone
DStudio – local DeepSeek V4 with a design studio, reachable from your phone
The Special Token `<Think>` Problem/Bug of Latest DeepSeek LLM
…Our evaluation of nine configurations across five model families -- GPT-5.4, Claude Sonnet 4.6, Gemini 2.5 Flash, DeepSeek V4 Flash, Qwen3-32B and GPT-5 mini -- reveals a steep…
…What’s most remarkable is that the attack was entirely crafted and executed by the open-source model DeepSeek-V3. The model crafted the opening gambit then responded to replies in ways…
…Up until now, users relying on cloud-based open models could use DeepSeek V4, Pro, GLM-5.2, Kimi K2.6, and Kimi-K2.7-Code. Today, the company added Kimi K3…
…In this paper, we moved beyond static, single-turn prompts to analyze multi-turn adversarial dialogues across distilled models like DeepSeek-R1-7B, Phi-4-Mini, and Qwen-4B-Thinking. By introducing…
…Library. Explore Toolkits | Stand-Alone Tools Latest Tech Insights Does DeepSeek* Solve the Small Scale Model Performance Puzzle? Learn how the DeepSeek-R1 distilled reasoning model performs and see how it works…
…It also generalizes across models and harnesses, improving DeepSeek-V4-Flash (OpenCode) from 12.13 to 19.58. View arXiv page View PDF GitHub 14 Add to collection Community AutoTrainess: Teaching Language…