Archives de catégorie : Non classé

Unlocking AI Inference Performance with Intel® Xeon® 6 Processors with Priority Core Turbo

As Large Language Models continue to scale, the CPU host node in CPU-GPU AI accelerated systems has become an important performance component.

Publié dans Non classé | Commentaires fermés sur Unlocking AI Inference Performance with Intel® Xeon® 6 Processors with Priority Core Turbo

Starting with Production in Mind: A Blueprint for Affordable Enterprise-Grade RAG on VMware Tanzu

Enterprise AI on CPUs: Intel & T‑Systems prove RAG runs cost‑effectively without GPUs.

Publié dans Non classé | Commentaires fermés sur Starting with Production in Mind: A Blueprint for Affordable Enterprise-Grade RAG on VMware Tanzu

Agentic Code Execution: A Leaner Way to Build AI Agents with Open Models

The landscape of AI agents is shifting from simple « ask-and-receive » interactions toward complex, multi-step reasoning.

Publié dans Non classé | Commentaires fermés sur Agentic Code Execution: A Leaner Way to Build AI Agents with Open Models

Intel® Xeon® 6 Processors: The Ultimate Host CPU Solution for AI-Accelerated Systems and Agentic AI

When it comes to choosing the right foundation for your AI infrastructure, Intel Xeon 6 processors stand out as the clear winner.

Publié dans Non classé | Commentaires fermés sur Intel® Xeon® 6 Processors: The Ultimate Host CPU Solution for AI-Accelerated Systems and Agentic AI

Running the AI Factory: How Enterprises Operationalize AI Placement at Scale

Part 3 of our CPU:GPU blog series now moves from metaphor to operations.

Publié dans Non classé | Commentaires fermés sur Running the AI Factory: How Enterprises Operationalize AI Placement at Scale

Intel® Xeon® 6 Processors and Intel® AMX Deliver More Concurrent Users with NVIDIA HGX B200 Systems

This blog introduces a heterogeneous architecture that co-runs vLLMs on both CPUs and GPUs to improve overall system efficiency.

Publié dans Non classé | Commentaires fermés sur Intel® Xeon® 6 Processors and Intel® AMX Deliver More Concurrent Users with NVIDIA HGX B200 Systems

Speed-up JAX LLM Training on Intel® Xeon® 6 CPU: Activation Offloading on Heterogeneous Systems

JAX-based Activation Offloading on Intel® Xeon® 6 with P-cores systems offers an effective alternative to activation recomputation, repurposing the CPU’s large DDR5 host memory as a live activation store and leveraging XLA’s asynchronous compute–communication overlap to avoid throughput loss.

Publié dans Non classé | Commentaires fermés sur Speed-up JAX LLM Training on Intel® Xeon® 6 CPU: Activation Offloading on Heterogeneous Systems

Unleash the Power of Intel® Xeon® 6 Processors with P-cores as AI Host CPU with Priority Core Turbo

By enabling priority cores to exceed standard all-core turbo limits and approach higher turbo frequency levels, PCT accelerates essential operations, reduces processing delays, and improves overall system responsiveness.

Publié dans Non classé | Commentaires fermés sur Unleash the Power of Intel® Xeon® 6 Processors with P-cores as AI Host CPU with Priority Core Turbo

Trust in the Age of Sovereign AI with Intel, Nvidia, and Red Hat

Confidential Computing is emerging as a foundational technology for Sovereign AI.

Publié dans Non classé | Commentaires fermés sur Trust in the Age of Sovereign AI with Intel, Nvidia, and Red Hat

Unlock Cost-Effective Enterprise AI with Lenovo ThinkSystem SR650 V4 with Intel® Xeon® 6 CPUs

We explore the scalability and performance of agentic document summarization running on a Red Hat OpenShift cluster built on Lenovo ThinkSystem SR650 V4 servers powered by Intel Xeon 6 processors.

Publié dans Non classé | Commentaires fermés sur Unlock Cost-Effective Enterprise AI with Lenovo ThinkSystem SR650 V4 with Intel® Xeon® 6 CPUs