Brand: Intel | Category: GPUs
SKU: ALTUS-G3-8P | Part #: ALTUS-G3-8P | MPN: ALTUS-G3-8P
Contact for Pricing — Request a Quote
The Penguin Computing Altus G3-8P is a high-density AI accelerator server built around eight Intel Gaudi 3 OAM (Open Accelerator Module) processors, designed to address the most demanding large-scale AI training and inference workloads in enterprise and hyperscale datacenter environments. Intel Gaudi 3 delivers a substantial generational leap over its predecessor, featuring a 64 MB on-chip SRAM, 128 GB HBM2e memory per OAM module (across the full 8-OAM configuration), and 24 Tensor Processor Cores per die alongside Matrix Multiplication Engines optimized for BF16 and FP8 precision. The architecture integrates 24 x 200 Gb/s RDMA-capable Ethernet ports per OAM for scale-out fabric connectivity, enabling high-bandwidth, low-latency communication across multi-node AI clusters without requiring proprietary interconnect hardware.
The Altus G3-8P platform from Penguin Computing is engineered for open-standards AI infrastructure, supporting the OCP Open Accelerator Infrastructure (OAI) form factor and offering deep integration with Intel's Gaudi software stack, including the Intel Gaudi PyTorch bridge and SynapseAI SDK. This allows enterprises to run leading AI frameworks—including PyTorch and TensorFlow—with optimized kernel libraries and model parallelism strategies across all eight accelerators. The server is positioned as a turnkey, rack-ready system that pairs the Gaudi 3 OAM modules with a validated host CPU platform, high-capacity system memory, and NVMe storage to deliver a complete, production-ready AI compute node.
Targeted at enterprise IT teams, national AI research institutions, and cloud service operators across the UAE, GCC, EMEA, and APAC regions, the Altus G3-8P represents a compelling alternative in the AI accelerator server segment. Its reliance on standard 200 GbE networking reduces fabric complexity, and the open-ecosystem software stack lowers long-term dependency risk. The system is suitable for generative AI model training, large language model (LLM) fine-tuning, computer vision pipelines, and high-throughput inference serving at scale.
| Manufacturer | Intel |
| Product Line | Penguin Computing Altus |
| Manufacturer Part Number | ALTUS-G3-8P |
| AI Accelerator | Intel Gaudi 3 OAM |
| Number of Accelerators | 8 x Intel Gaudi 3 OAM modules |
| Accelerator On-Chip SRAM (per OAM) | 64 MB |
| Accelerator Memory Type | HBM2e |
| Accelerator Memory (per OAM) | 128 GB HBM2e |
| Total Accelerator Memory (8 OAM) | 1 TB HBM2e |
| Tensor Processor Cores (per OAM) | 24 |
| Supported Precisions | FP8, BF16, FP16, FP32, INT8 |
| Scale-Out Networking (per OAM) | 24 x 200 Gb/s Ethernet (RDMA-capable) |
| Total Scale-Out Ethernet Ports | 192 x 200 Gb/s (across 8 OAM modules) |
| Accelerator Form Factor | OCP Open Accelerator Module (OAM) |
| Software Stack | Intel SynapseAI SDK, Intel Gaudi PyTorch integration |
| Supported Frameworks | PyTorch, TensorFlow |
| Server Form Factor | 4U Rack-Mount |
| OAM Standard | OCP Open Accelerator Infrastructure (OAI) |
Available from Omnixon Global. Submit an RFQ and our team will confirm configuration and availability for your order.
| Brand | Intel |
| Category | GPUs |
| SKU | ALTUS-G3-8P |
| Part Number | ALTUS-G3-8P |
| Condition | New |
| Product Line | Penguin Computing Altus |
| Manufacturer Part Number | ALTUS-G3-8P |
| AI Accelerator | Intel Gaudi 3 OAM |
| Number of Accelerators | 8 x Intel Gaudi 3 OAM modules |
| Accelerator On-Chip SRAM (per OAM) | 64 MB |
| Accelerator Memory Type | HBM2e |
| Accelerator Memory (per OAM) | 128 GB HBM2e |
| Total Accelerator Memory (8 OAM) | 1 TB HBM2e |
| Tensor Processor Cores (per OAM) | 24 |
| Supported Precisions | FP8, BF16, FP16, FP32, INT8 |
| Scale-Out Networking (per OAM) | 24 x 200 Gb/s Ethernet (RDMA-capable) |
| Total Scale-Out Ethernet Ports | 192 x 200 Gb/s (across 8 OAM modules) |
| Accelerator Form Factor | OCP Open Accelerator Module (OAM) |
| Software Stack | Intel SynapseAI SDK, Intel Gaudi PyTorch integration |
| Supported Frameworks | PyTorch, TensorFlow |
| Server Form Factor | 4U Rack-Mount |
| OAM Standard | OCP Open Accelerator Infrastructure (OAI) |
Reference servers include Dell PowerEdge XE9680 / XE9712, HPE Cray XD670, Lenovo ThinkSystem SR685a / SR675 V3, Supermicro AS-A21GE / SYS-821GE, Gigabyte G593 / G894, ASUS ESC. Share your target platform in the RFQ and we will confirm chassis-to-GPU compatibility and recommended NIC pairing.
Highly model-dependent. L40S / RTX-class: typically 3-6 weeks. H100/H200/B200 in SXM form factor: 12-16 weeks for whole-platform allocations. We quote genuine-channel ETAs only — no grey-market promises.
Yes — Omnixon stocks the full NVIDIA networking lineup (Quantum-2 / Quantum-X InfiniBand, Spectrum-X Ethernet, ConnectX NICs, BlueField DPUs) so we can quote a complete training-cluster BOM, not just the GPUs.
Yes. We hold genuine-channels for NVIDIA AI Enterprise software subscriptions. Add it to your RFQ and we quote node-aligned licensing along with the hardware.