Search

Showing top 7 results for "Model performance/limits"

Filtered by topic: OpenAI Clear ✕

tomshardware.com › tech-industry › artificial-intelligence

China's 2.8-trillion-parameter Kimi K3 beats Claude Fable 5 in Frontend Code Arena benchmark—  Moonshot AI delivers largest open-weight AI model ever, as China works around U.S. compute limits

… Moonshot said K3 still sits behind Anthropic's Claude Fable 5 and OpenAI's GPT 5.6 Sol on overall performance, but it outperformed every other model in the company's evaluation suite, including Claude Opus 4.8 and GPT 5.5, across coding and agentic benchmarks . …

Jul 17, 2026 · Luke James
tomshardware.com › tech-industry › artificial-intelligence

OpenAI's GPT-5.6 Sol and unreleased AI models break out of testing environment in 'unprecedented cybersecurity incident' — rogue agents hacked HuggingFace's production servers with 'thousands of individual actions across a swarm of short-lived sandboxes'

… For its part, OpenAI says it's going to add controls to the bot "at the cost of research velocity," and that the incident "points to the need to further strengthen our model’s alignment, cyber protections during evaluation time, and monitoring during internal testing" — a statement that might be an… …

Jul 22, 2026 · Bruno Ferreira
tomshardware.com › tech-industry › artificial-intelligence

Broadcom and OpenAI unveil custom-built Jalapeño inference processor — OpenAI's first chip is a massive reticle-sized ASIC built in an ultra-fast nine-month development cycle

OpenAI and Broadcom have introduced Jalapeño, a custom-built inference processor designed specifically for modern large language models and future agentic AI workloads, which is designed to deliver performance per watt they claim is higher than today's leading-edge hardware. …

Jun 24, 2026 · Anton Shilov
2 sources covering this — show 1 more