<?xml version="1.0" encoding="UTF-8"?>
<rss version="2.0" xmlns:atom="http://www.w3.org/2005/Atom">
<channel>
  <title>Eurokommerz Engineering Blog</title>
  <link>https://eurokommerz.at/blog/</link>
  <atom:link href="https://eurokommerz.at/blog/rss.xml" rel="self" type="application/rss+xml"/>
  <description>Deep dives on AI infrastructure: GPUs, memory sizing, virtualisation and the parts vendor slides skip.</description>
  <language>en</language>
  <lastBuildDate>Mon, 14 Sep 2026 09:00:00 +0200</lastBuildDate>
  <item>
    <title>VMware Renewal 2026: Standard, VVF or VCF</title>
    <link>https://eurokommerz.at/blog/vmware-renewal-standard-vvf-vcf-which-tier/</link>
    <guid isPermaLink="true">https://eurokommerz.at/blog/vmware-renewal-standard-vvf-vcf-which-tier/</guid>
    <pubDate>Mon, 14 Sep 2026 09:00:00 +0200</pubDate>
    <description>Standard and Enterprise Plus stop at vSphere 8 U3, 9.x lives only in VVF and VCF: features per tier, per-core rules, vSAN entitlements, and the renewal traps.</description>
  </item>
  <item>
    <title>RTX PRO 6000 vs 5000 vs 4500 Blackwell: Which Card</title>
    <link>https://eurokommerz.at/blog/rtx-pro-6000-vs-5000-vs-4500-blackwell/</link>
    <guid isPermaLink="true">https://eurokommerz.at/blog/rtx-pro-6000-vs-5000-vs-4500-blackwell/</guid>
    <pubDate>Mon, 14 Sep 2026 09:00:00 +0200</pubDate>
    <description>96, 72, 48, 32 or 24 GB: memory, bandwidth, MIG and vGPU for each RTX PRO Blackwell card, which LLM fits where, and why CAD does not need the big one.</description>
  </item>
  <item>
    <title>NVIDIA AI Enterprise Licensing Explained</title>
    <link>https://eurokommerz.at/blog/nvidia-ai-enterprise-licensing-explained/</link>
    <guid isPermaLink="true">https://eurokommerz.at/blog/nvidia-ai-enterprise-licensing-explained/</guid>
    <pubDate>Mon, 14 Sep 2026 09:00:00 +0200</pubDate>
    <description>Per GPU, not per VM: how NVIDIA AI Enterprise is counted, what ships with H100 NVL and H200 NVL, when vGPU compute needs it, and what runs without it.</description>
  </item>
  <item>
    <title>Immutable Backup: What It Really Protects Against</title>
    <link>https://eurokommerz.at/blog/immutable-backup-what-it-protects-against/</link>
    <guid isPermaLink="true">https://eurokommerz.at/blog/immutable-backup-what-it-protects-against/</guid>
    <pubDate>Mon, 14 Sep 2026 09:00:00 +0200</pubDate>
    <description>Hardened repositories and S3 Object Lock explained, the attacks immutability stops and the ones it does not, and how to size the lock window against dwell time.</description>
  </item>
  <item>
    <title>How Many Users One RTX PRO 6000 Can Serve</title>
    <link>https://eurokommerz.at/blog/how-many-users-per-rtx-pro-6000/</link>
    <guid isPermaLink="true">https://eurokommerz.at/blog/how-many-users-per-rtx-pro-6000/</guid>
    <pubDate>Mon, 14 Sep 2026 09:00:00 +0200</pubDate>
    <description>NVIDIA publishes no tokens/s for this card. The KV cache arithmetic, the 1,597 GB/s server figure and the vLLM log line that prints your own concurrency.</description>
  </item>
  <item>
    <title>H200 NVL in an Existing Server: Checklist</title>
    <link>https://eurokommerz.at/blog/h200-nvl-in-an-existing-server-checklist/</link>
    <guid isPermaLink="true">https://eurokommerz.at/blog/h200-nvl-in-an-existing-server-checklist/</guid>
    <pubDate>Mon, 14 Sep 2026 09:00:00 +0200</pubDate>
    <description>Slot, sense pin, cooling, support list, NVLink, BIOS and licence: the seven gates that decide whether 141 GB of HBM3e can go into the server you already have.</description>
  </item>
  <item>
    <title>First AI Project: DGX Spark, RTX PRO or H200</title>
    <link>https://eurokommerz.at/blog/which-gpu-for-a-first-ai-project/</link>
    <guid isPermaLink="true">https://eurokommerz.at/blog/which-gpu-for-a-first-ai-project/</guid>
    <pubDate>Mon, 14 Sep 2026 09:00:00 +0200</pubDate>
    <description>Model size, users at once, where the box lives, what follows the pilot: 273 GB/s against 1,597 GB/s and 4.8 TB/s, and which of the three a first project needs.</description>
  </item>
  <item>
    <title>Fine-Tuning vs RAG: Which One, and What It Costs</title>
    <link>https://eurokommerz.at/blog/fine-tuning-vs-rag-which-one/</link>
    <guid isPermaLink="true">https://eurokommerz.at/blog/fine-tuning-vs-rag-which-one/</guid>
    <pubDate>Mon, 14 Sep 2026 09:00:00 +0200</pubDate>
    <description>RAG changes the input, fine-tuning changes the weights. What each one fixes, what full tuning, LoRA and QLoRA cost in VRAM, and what the papers really found.</description>
  </item>
  <item>
    <title>DGX Spark 128 GB: What Fits, What Tunes</title>
    <link>https://eurokommerz.at/blog/what-fits-on-a-dgx-spark-128gb/</link>
    <guid isPermaLink="true">https://eurokommerz.at/blog/what-fits-on-a-dgx-spark-128gb/</guid>
    <pubDate>Mon, 14 Sep 2026 09:00:00 +0200</pubDate>
    <description>No VRAM figure, one shared pool: what the 128 GB really holds at BF16, FP8 and FP4, why 200 billion means 4-bit, and what one unit can actually fine-tune.</description>
  </item>
  <item>
    <title>DGX B300 vs 8× RTX PRO 6000 Server: NVLink or PCIe</title>
    <link>https://eurokommerz.at/blog/dgx-b300-vs-8x-rtx-pro-6000-server/</link>
    <guid isPermaLink="true">https://eurokommerz.at/blog/dgx-b300-vs-8x-rtx-pro-6000-server/</guid>
    <pubDate>Mon, 14 Sep 2026 09:00:00 +0200</pubDate>
    <description>2.1 TB of HBM3e with NVLink at 14 kW, or 768 GB of GDDR7 on PCIe in 4U: memory arithmetic, MLPerf v6.0 results, facility demands and the workloads each fits.</description>
  </item>
  <item>
    <title>Backup Is Not Disaster Recovery</title>
    <link>https://eurokommerz.at/blog/backup-is-not-disaster-recovery/</link>
    <guid isPermaLink="true">https://eurokommerz.at/blog/backup-is-not-disaster-recovery/</guid>
    <pubDate>Mon, 14 Sep 2026 09:00:00 +0200</pubDate>
    <description>A successful backup job is evidence about RPO, not RTO. What a restore leaves behind, how long 82 TB really takes, and what proven recovery has to show.</description>
  </item>
  <item>
    <title>GPU Server for a 70B LLM: Worked Sizing Example</title>
    <link>https://eurokommerz.at/blog/gpu-server-for-70b-llm-worked-example/</link>
    <guid isPermaLink="true">https://eurokommerz.at/blog/gpu-server-for-70b-llm-worked-example/</guid>
    <pubDate>Sun, 13 Sep 2026 09:00:00 +0200</pubDate>
    <description>Llama 3.3 70B, 30 users, 8k context: the memory arithmetic, why one 96 GB card only just works in FP4 and not in FP8, the H200 NVL option, the server around it.</description>
  </item>
  <item>
    <title>Liquid Cooling for GPU Racks: 50 to 120 kW</title>
    <link>https://eurokommerz.at/blog/liquid-cooling-for-gpu-racks/</link>
    <guid isPermaLink="true">https://eurokommerz.at/blog/liquid-cooling-for-gpu-racks/</guid>
    <pubDate>Sat, 12 Sep 2026 09:00:00 +0200</pubDate>
    <description>Where air stops (20 to 50 kW per rack), what rear doors, cold plates and immersion each carry, ASHRAE water classes, flow rates, CDU sizes, retrofit realities.</description>
  </item>
  <item>
    <title>How Many GPUs Fit in One Server: Lanes, Power, Air</title>
    <link>https://eurokommerz.at/blog/how-many-gpus-fit-in-one-server/</link>
    <guid isPermaLink="true">https://eurokommerz.at/blog/how-many-gpus-fit-in-one-server/</guid>
    <pubDate>Fri, 11 Sep 2026 09:00:00 +0200</pubDate>
    <description>Why a 2U takes four 350 W cards but two 600 W ones, where PCIe switches come in, what an eight-GPU server draws, and the number that stops most orders.</description>
  </item>
  <item>
    <title>DGX Spark: Two Nodes, 256 GB, and Beyond LLMs</title>
    <link>https://eurokommerz.at/blog/dgx-spark-two-nodes-and-beyond-llms/</link>
    <guid isPermaLink="true">https://eurokommerz.at/blog/dgx-spark-two-nodes-and-beyond-llms/</guid>
    <pubDate>Thu, 10 Sep 2026 09:00:00 +0200</pubDate>
    <description>How two units link, what NVIDIA supports up to four, measured two-node token rates, image and video generation, fine-tuning, and the honest answer on gaming.</description>
  </item>
  <item>
    <title>RTX PRO 6000 vs H200 NVL: 96 GB or 141 GB for LLMs</title>
    <link>https://eurokommerz.at/blog/rtx-pro-6000-vs-h200-nvl-llm-inference/</link>
    <guid isPermaLink="true">https://eurokommerz.at/blog/rtx-pro-6000-vs-h200-nvl-llm-inference/</guid>
    <pubDate>Wed, 09 Sep 2026 09:00:00 +0200</pubDate>
    <description>96 GB GDDR7 at 1.6 TB/s with FP4, or 141 GB HBM3e at 4.8 TB/s with NVLink. Model footprints, the bandwidth arithmetic behind tokens per second, MLPerf results.</description>
  </item>
  <item>
    <title>L40S vs RTX PRO 6000 Server: 48 GB vs 96 GB</title>
    <link>https://eurokommerz.at/blog/l40s-vs-rtx-pro-6000-server-edition/</link>
    <guid isPermaLink="true">https://eurokommerz.at/blog/l40s-vs-rtx-pro-6000-server-edition/</guid>
    <pubDate>Tue, 08 Sep 2026 09:00:00 +0200</pubDate>
    <description>48 GB at 350 W or 96 GB at up to 600 W: memory, bandwidth, MIG, vGPU seats, server fit, and the cases where the older L40S is still the right order.</description>
  </item>
  <item>
    <title>RAG on Company Data: Retrieval Fails First</title>
    <link>https://eurokommerz.at/blog/rag-on-company-data/</link>
    <guid isPermaLink="true">https://eurokommerz.at/blog/rag-on-company-data/</guid>
    <pubDate>Fri, 04 Sep 2026 09:00:00 +0200</pubDate>
    <description>Almost every RAG failure is a retrieval failure, not a model failure. Permissions at retrieval time, chunking and metadata, the evaluation set nobody builds, and the index that quietly goes stale.</description>
  </item>
  <item>
    <title>Private LLM vs Cloud API: Where the Break-Even Is</title>
    <link>https://eurokommerz.at/blog/private-llm-vs-cloud-api/</link>
    <guid isPermaLink="true">https://eurokommerz.at/blog/private-llm-vs-cloud-api/</guid>
    <pubDate>Thu, 03 Sep 2026 09:00:00 +0200</pubDate>
    <description>The break-even is a token-volume question. The four cost lines nobody counts, the arithmetic on your own numbers, and when residency decides first.</description>
  </item>
  <item>
    <title>vSphere Waste: Oversized VMs, Snapshots, Headroom</title>
    <link>https://eurokommerz.at/blog/vsphere-where-the-waste-is/</link>
    <guid isPermaLink="true">https://eurokommerz.at/blog/vsphere-where-the-waste-is/</guid>
    <pubDate>Wed, 02 Sep 2026 09:00:00 +0200</pubDate>
    <description>Oversized VMs, zombie workloads, forgotten snapshots and stale cluster headroom: the four places capacity goes, how to measure them honestly, and why more vCPU often makes a VM slower.</description>
  </item>
  <item>
    <title>RPO vs RTO: The Difference and How to Set Them</title>
    <link>https://eurokommerz.at/blog/rpo-rto-how-to-choose/</link>
    <guid isPermaLink="true">https://eurokommerz.at/blog/rpo-rto-how-to-choose/</guid>
    <pubDate>Wed, 02 Sep 2026 09:00:00 +0200</pubDate>
    <description>RPO is how much data you can lose, RTO how long you can be down. Why cost climbs steeply near zero, how to tier systems, the dependency order that breaks plans.</description>
  </item>
  <item>
    <title>VMware Licensing After Broadcom: 16-Core Minimum</title>
    <link>https://eurokommerz.at/blog/vmware-licensing-after-broadcom/</link>
    <guid isPermaLink="true">https://eurokommerz.at/blog/vmware-licensing-after-broadcom/</guid>
    <pubDate>Tue, 01 Sep 2026 09:00:00 +0200</pubDate>
    <description>How VCF and vSphere Foundation cores are counted, what the 16-core minimum does to a small cluster, what the 72-core story really was, and where the number gets larger than it needs to be.</description>
  </item>
  <item>
    <title>GPU Rack Power and Cooling: kW, Airflow, Liquid</title>
    <link>https://eurokommerz.at/blog/gpu-rack-power-and-cooling/</link>
    <guid isPermaLink="true">https://eurokommerz.at/blog/gpu-rack-power-and-cooling/</guid>
    <pubDate>Mon, 24 Aug 2026 09:00:00 +0200</pubDate>
    <description>Rack density in kW, the airflow arithmetic, the point where air cooling stops and liquid starts, and the four checks that decide whether a delivery switches on.</description>
  </item>
  <item>
    <title>NVIDIA DGX Spark: up to 55 tok/s, 273 GB/s, 140 W</title>
    <link>https://eurokommerz.at/blog/nvidia-dgx-spark-real-benchmarks/</link>
    <guid isPermaLink="true">https://eurokommerz.at/blog/nvidia-dgx-spark-real-benchmarks/</guid>
    <pubDate>Sun, 09 Aug 2026 09:00:00 +0200</pubDate>
    <description>NVIDIA’s own figure for gpt-oss 120B: 55 tok/s generation, 1,725 tok/s prefill. The 273 GB/s ceiling, 128 GB unified memory, 140 W TDP, two-node pairing.</description>
  </item>
  <item>
    <title>Nutanix NX-9151, NX-3155, NX-8150G: GPUs per Node</title>
    <link>https://eurokommerz.at/blog/nutanix-gpu-ai-cluster/</link>
    <guid isPermaLink="true">https://eurokommerz.at/blog/nutanix-gpu-ai-cluster/</guid>
    <pubDate>Wed, 05 Aug 2026 09:00:00 +0200</pubDate>
    <description>Which NX nodes (G9 and G10) take L40S, H100 NVL, A16 and L4, how many cards each holds, what AHV supports, and why power and airflow, not case size, block the next card.</description>
  </item>
  <item>
    <title>RTX PRO 6000 Blackwell vs RTX 6000 Ada: 96 vs 48 GB</title>
    <link>https://eurokommerz.at/blog/rtx-pro-6000-blackwell-vs-rtx-6000-ada/</link>
    <guid isPermaLink="true">https://eurokommerz.at/blog/rtx-pro-6000-blackwell-vs-rtx-6000-ada/</guid>
    <pubDate>Tue, 28 Jul 2026 09:00:00 +0200</pubDate>
    <description>96 GB GDDR7 vs 48 GB GDDR6, 1,792 vs 960 GB/s, FP4 and MIG that Ada never had, and the cases where RTX 6000 Ada still does the job.</description>
  </item>
  <item>
    <title>RTX PRO 6000: Workstation vs Max-Q vs Server</title>
    <link>https://eurokommerz.at/blog/rtx-pro-6000-workstation-max-q-server-comparison/</link>
    <guid isPermaLink="true">https://eurokommerz.at/blog/rtx-pro-6000-workstation-max-q-server-comparison/</guid>
    <pubDate>Mon, 20 Jul 2026 09:00:00 +0200</pubDate>
    <description>Same 96 GB silicon, three power envelopes: 600 W, 300 W and 400–600 W. Which one takes vGPU, how many VMs per card, and which belongs in a server.</description>
  </item>
  <item>
    <title>MIG vs vGPU: VMs per RTX PRO 6000, 5000 and L40S</title>
    <link>https://eurokommerz.at/blog/mig-vgpu-how-many-vms-per-gpu/</link>
    <guid isPermaLink="true">https://eurokommerz.at/blog/mig-vgpu-how-many-vms-per-gpu/</guid>
    <pubDate>Fri, 10 Jul 2026 09:00:00 +0200</pubDate>
    <description>MIG splits the silicon, vGPU splits time and needs a per-user licence. Maximum instances and smallest slice for each card, and where the setup breaks.</description>
  </item>
  <item>
    <title>H200 NVL vs H100 NVL: 141 GB vs 94 GB</title>
    <link>https://eurokommerz.at/blog/h200-nvl-vs-h100-nvl/</link>
    <guid isPermaLink="true">https://eurokommerz.at/blog/h200-nvl-vs-h100-nvl/</guid>
    <pubDate>Sun, 28 Jun 2026 09:00:00 +0200</pubDate>
    <description>Identical compute silicon, all the difference in memory: 141 vs 94 GB and 4.8 vs 3.9 TB/s. Where the gain reaches 3.4×, and where it is exactly zero.</description>
  </item>
  <item>
    <title>VRAM for an LLM: Sizing 7B, 13B and 70B Models</title>
    <link>https://eurokommerz.at/blog/how-much-vram-llm-needs/</link>
    <guid isPermaLink="true">https://eurokommerz.at/blog/how-much-vram-llm-needs/</guid>
    <pubDate>Mon, 15 Jun 2026 09:00:00 +0200</pubDate>
    <description>Weights, KV cache, CUDA-graph overhead and concurrent users: the working formulas and sizing tables for choosing a GPU for local LLM inference.</description>
  </item>
</channel>
</rss>
