{"product_id":"inference-35b-rtx4090-ai-server","title":"Inference 35B RTX4090 AI Server","description":"\u003cdiv style=\"background:#dc2626;color:#fff;padding:20px;border-radius:8px;margin-bottom:24px;border-left:6px solid #991b1b\"\u003e\n\u003cp style=\"font-size:18px;font-weight:700;margin:0 0 8px 0\"\u003eThis product listing is kept for reference only.\u003c\/p\u003e\n\u003cp style=\"margin:0\"\u003eThis server has been replaced by the new \u003cstrong\u003eKentino AI product line\u003c\/strong\u003e. For the current equivalent or an upgraded configuration, please see \u003ca href=\"\/en\/collections\/ai-servers\" style=\"color:#fca5a5;text-decoration:underline\"\u003eour AI Servers collection\u003c\/a\u003e.\u003c\/p\u003e\n\u003cp style=\"margin:8px 0 0 0\"\u003e\u003cstrong\u003eRecommended replacement:\u003c\/strong\u003e \u003ca href=\"\/en\/collections\/ai-servers\" style=\"color:#fca5a5;text-decoration:underline\"\u003eKentino AI 96 Rome 4090 2644TOPS\u003c\/a\u003e (4x RTX 4090, same platform, updated build)\u003c\/p\u003e\n\u003c\/div\u003e\n\u003ch2 class=\"font-600 text-xl font-bold\" level=\"2\"\u003eSpecifications\u003c\/h2\u003e\n\u003cul class=\"-mt-1 list-disc space-y-2 pl-8\" depth=\"0\"\u003e\n\u003cli class=\"whitespace-normal break-words\" index=\"0\"\u003e\n\u003cstrong\u003eGPU:\u003c\/strong\u003e 4x NVIDIA RTX 4090 (96 GB VRAM total)\u003c\/li\u003e\n\u003cli class=\"whitespace-normal break-words\" index=\"1\"\u003e\n\u003cstrong\u003eMotherboard:\u003c\/strong\u003e ASRock Rack ROMED8-2T\u003c\/li\u003e\n\u003cli class=\"whitespace-normal break-words\" index=\"2\"\u003e\n\u003cstrong\u003eCPU:\u003c\/strong\u003e AMD EPYC 7542\u003c\/li\u003e\n\u003cli class=\"whitespace-normal break-words\" index=\"3\"\u003e\n\u003cstrong\u003eRAM:\u003c\/strong\u003e 256GB A-Tech DDR4-2666 ECC REG RDIMM (8 x 32GB)\u003c\/li\u003e\n\u003cli class=\"whitespace-normal break-words\" index=\"4\"\u003e\n\u003cstrong\u003eGPU-Motherboard Connection:\u003c\/strong\u003e RYSER PCIe 4.0 x16 Cable\u003c\/li\u003e\n\u003cli class=\"whitespace-normal break-words\" index=\"5\"\u003e\n\u003cstrong\u003ePower Supply:\u003c\/strong\u003e 2x LL2000FC 4 Kw\u003c\/li\u003e\n\u003cli class=\"whitespace-normal break-words\" index=\"6\"\u003e\n\u003cstrong\u003eCase:\u003c\/strong\u003e 24U Rack Mount\u003c\/li\u003e\n\u003cli class=\"whitespace-normal break-words\" index=\"7\"\u003e\n\u003cstrong\u003eStorage:\u003c\/strong\u003e\n\u003cul class=\"-mt-1 list-disc space-y-2 pl-8\" depth=\"1\"\u003e\n\u003cli class=\"whitespace-normal break-words\" index=\"0\"\u003e2TB NVMe SSD\u003c\/li\u003e\n\u003cli class=\"whitespace-normal break-words\" index=\"1\"\u003e500GB SATA Drive\u003c\/li\u003e\n\u003c\/ul\u003e\n\u003c\/li\u003e\n\u003c\/ul\u003e\n\u003ch2 class=\"font-600 text-xl font-bold\" level=\"2\"\u003eKey Features\u003c\/h2\u003e\n\u003col class=\"-mt-1 list-decimal space-y-2 pl-8\" depth=\"0\"\u003e\n\u003cli class=\"whitespace-normal break-words\" index=\"0\"\u003eOptimized for AI Inference: Equipped with 4 NVIDIA RTX 4090 GPUs, providing a total of 96 GB VRAM, specifically configured for high-performance AI inference tasks, including large language models up to 70B parameters.\u003c\/li\u003e\n\u003cli class=\"whitespace-normal break-words\" index=\"1\"\u003eServer-Grade Components: Features the reliable ASRock Rack ROMED8-2T motherboard and a powerful AMD EPYC 7542 CPU for exceptional processing capabilities.\u003c\/li\u003e\n\u003cli class=\"whitespace-normal break-words\" index=\"2\"\u003eHigh-Speed Memory: 256GB of A-Tech DDR4-2666 ECC REG RDIMM ensures reliable and efficient data processing for complex AI workloads.\u003c\/li\u003e\n\u003cli class=\"whitespace-normal break-words\" index=\"3\"\u003eFast GPU Integration: Utilizes the RYSER PCIe 4.0 x16 cable for rapid, full-bandwidth connection between the GPUs and the motherboard, maximizing inference performance.\u003c\/li\u003e\n\u003cli class=\"whitespace-normal break-words\" index=\"4\"\u003eRobust Power Supply: An AX1600i 1500W unit provides stable and ample power delivery to support the high-performance components under intensive inference loads.\u003c\/li\u003e\n\u003cli class=\"whitespace-normal break-words\" index=\"5\"\u003eEfficient Storage: Comes with a fast 2TB NVMe SSD for quick data access and an additional 500GB SATA drive for extra capacity.\u003c\/li\u003e\n\u003cli class=\"whitespace-normal break-words\" index=\"6\"\u003eProfessional-Grade Cooling: Housed in a spacious 24U rack mount case, ensuring optimal thermal management for sustained high-performance operation.\u003c\/li\u003e\n\u003cli class=\"whitespace-normal break-words\" index=\"7\"\u003eInference-Focused Design: Optimized for running large AI models efficiently, making it ideal for organizations deploying AI services at scale.\u003c\/li\u003e\n\u003c\/ol\u003e\n\u003ch2 class=\"font-600 text-xl font-bold\" level=\"2\"\u003eIdeal Use Cases\u003c\/h2\u003e\n\u003cul class=\"-mt-1 list-disc space-y-2 pl-8\" depth=\"0\"\u003e\n\u003cli class=\"whitespace-normal break-words\" index=\"0\"\u003eLarge Language Model Inference (up to 70B parameters)\u003c\/li\u003e\n\u003cli class=\"whitespace-normal break-words\" index=\"1\"\u003eReal-time AI-powered Applications\u003c\/li\u003e\n\u003cli class=\"whitespace-normal break-words\" index=\"2\"\u003eNatural Language Processing Services\u003c\/li\u003e\n\u003cli class=\"whitespace-normal break-words\" index=\"3\"\u003eComputer Vision and Image Recognition\u003c\/li\u003e\n\u003cli class=\"whitespace-normal break-words\" index=\"4\"\u003eAI-driven Customer Service and Chatbots\u003c\/li\u003e\n\u003cli class=\"whitespace-normal break-words\" index=\"5\"\u003eRecommendation Systems\u003c\/li\u003e\n\u003cli class=\"whitespace-normal break-words\" index=\"6\"\u003eFinancial Modeling and Predictions\u003c\/li\u003e\n\u003cli class=\"whitespace-normal break-words\" index=\"7\"\u003eScientific Data Analysis\u003c\/li\u003e\n\u003c\/ul\u003e\n\u003ch2 class=\"font-600 text-xl font-bold\" level=\"2\"\u003eSpecial Notes\u003c\/h2\u003e\n\u003cul class=\"-mt-1 list-disc space-y-2 pl-8\" depth=\"0\"\u003e\n\u003cli class=\"whitespace-normal break-words\" index=\"0\"\u003eRTX 4090 Advantage: Leveraging the latest NVIDIA RTX 4090 GPUs, this server offers exceptional performance for AI inference tasks, combining high compute power with advanced features like Tensor Cores.\u003c\/li\u003e\n\u003cli class=\"whitespace-normal break-words\" index=\"1\"\u003eOptimized for 70B Models: With 96 GB of total GPU VRAM, this system is specifically designed to handle large language models with up to 70 billion parameters, making it ideal for deploying state-of-the-art AI services.\u003c\/li\u003e\n\u003cli class=\"whitespace-normal break-words\" index=\"2\"\u003eInference Efficiency: The combination of RTX 4090 GPUs and the AMD EPYC CPU allows for highly efficient inference, enabling high throughput and low latency for AI applications.\u003c\/li\u003e\n\u003cli class=\"whitespace-normal break-words\" index=\"3\"\u003eScalable Solution: While optimized for 70B parameter models, this server can be easily integrated into larger clusters for even more demanding workloads or multi-model deployments.\u003c\/li\u003e\n\u003c\/ul\u003e\n\u003cp class=\"whitespace-pre-wrap break-words\"\u003eThe Inference 70B RTX4090 AI Server is a cutting-edge solution for organizations looking to deploy large AI models efficiently. It strikes an optimal balance between performance and cost, making it an excellent choice for businesses and research institutions that need to run complex AI models in production environments. Whether you're deploying language models, computer vision systems, or other AI applications, this server provides the power and reliability needed for seamless AI inference at scale.\u003c\/p\u003e\n\u003cp class=\"whitespace-pre-wrap break-words\"\u003eDelivery 2 - 6 weeks \u003c\/p\u003e","brand":"Kentino","offers":[{"title":"Default Title","offer_id":49061697913160,"sku":"","price":14909.0,"currency_code":"EUR","in_stock":true}],"thumbnail_url":"\/\/cdn.shopify.com\/s\/files\/1\/0843\/5479\/3800\/files\/inference-35b-rtx4090-ai-server.jpg?v=1786104430","url":"https:\/\/kentino.se\/en\/products\/inference-35b-rtx4090-ai-server","provider":"Kentino","version":"1.0","type":"link"}