Accelerate Microsoft Phi-4 Small Language Models
…Our preliminary benchmarking of Phi-4-mini and Phi-4-multimodal using PyTorch and Intel Extension for PyTorch on Intel Xeon 6 with MRDIMMs demonstrates that widely available Xeon processors are a…
Tracked topic
Get started with Intel Extension for PyTorch today and accelerate your PyTorch training and inference performance on 4th Gen Intel Xeon Scalable processors using Intel AMX. We encourage you to also check out and incorporate Intel’s other AI/ML Framework optimizations and end-to-end portfolio of tools into your AI workflow and learn about the unified, open, standards-based oneAPI programming model that forms the foundation of Intel’s AI Software Portfolio to help you prepare, build, deploy, and scale your AI solutions. For more details about the new 4th Gen Intel Xeon Scalable processors, visit
Accelerate PyTorch Training and Inference using Intel® AMX…Our preliminary benchmarking of Phi-4-mini and Phi-4-multimodal using PyTorch and Intel Extension for PyTorch on Intel Xeon 6 with MRDIMMs demonstrates that widely available Xeon processors are a…
…Discover application performance characterization, such as bandwidth sensitivity, instruction mix, and cache-line use, for Intel GPUs, multi-tile architectures, and VNNI and ISA support. Extensions for TensorFlow and extensions for PyTorch…
…Microsoft* "The Intel team's optimization of fMRI and PadChest models using Intel® Extension for PyTorch* and OpenVINO™ toolkit powered by oneAPI, leading to approximately 6x increase in performance, tailored for medical…
…Using Intel® Extension for PyTorch* and high-performance Intel® GPUs, the team designed and trained a stable diffusion control pipeline, enabling prompt-driven material creation for visualization and rendering tasks. Intel® Data…
By Intel® Neural Compressor is an open source Python* library designed to help quickly optimize inference solutions on popular deep learning frameworks (TensorFlow*, PyTorch*, ONNX* [Open Neural Network Exchange] runtime, and Apache…
By Intel® Neural Compressor is an open source Python* library designed to help quickly optimize inference solutions on popular deep learning frameworks (TensorFlow*, PyTorch, Open Neural Network Exchange [ONNX*] runtime, and MXNet…
…Building an End-to-End AI Solution Using PyTorch* How to Build an Interactive Chat-Generation Model Using DialoGPT and PyTorch* Accelerate PyTorch Training and Inference Performance Using Intel AMX See All…
…This retail use case shows an example of fine-tuning a YOLOX-PyTorch* model to manage shelf space and shelf inventory. This example takes a dataset of store shelf images as input…
…Seekr LLM fine-tuning using the Hugging Face Optimum Habana and Habana PyTorch APIs. Source: Seekr. Intel does not control or audit third-party data. You should consult other sources to evaluate…
…Notebook. How It Works Convert and optimize models trained using popular frameworks like TensorFlow* and PyTorch*. Deploy across a mix of Intel® hardware and environments, on-premise and on-device, in the…