Sony Pixel Power calrec Sony

Oracle Cloud Infrastructure Expands NVIDIA GPU-Accelerated Instances for AI, Digital Twins and More

31/07/2024

Enterprises are rapidly adopting generative AI, large language models (LLMs), advanced graphics and digital twins to increase operational efficiencies, reduce costs and drive innovation.

However, to adopt these technologies effectively, enterprises need access to state-of-the-art, full-stack accelerated computing platforms. To meet this demand, Oracle Cloud Infrastructure (OCI) today announced NVIDIA L40S GPU bare-metal instances available to order and the upcoming availability of a new virtual machine accelerated by a single NVIDIA H100 Tensor Core GPU. This new VM expands OCI's existing H100 portfolio, which includes an NVIDIA HGX H100 8-GPU bare-metal instance.

Paired with NVIDIA networking and running the NVIDIA software stack, these platforms deliver powerful performance and efficiency, enabling enterprises to advance generative AI.

NVIDIA L40S Now Available to Order on OCI The NVIDIA L40S is a universal data center GPU designed to deliver breakthrough multi-workload acceleration for generative AI, graphics and video applications. Equipped with fourth-generation Tensor Cores and support for the FP8 data format, the L40S GPU excels in training and fine-tuning small- to mid-size LLMs and in inference across a wide range of generative AI use cases.

For example, a single L40S GPU (FP8) can generate up to 1.4x more tokens per second than a single NVIDIA A100 Tensor Core GPU (FP16) for Llama 3 8B with NVIDIA TensorRT-LLM at an input and output sequence length of 128.

The L40S GPU also has best-in-class graphics and media acceleration. Its third-generation NVIDIA Ray Tracing Cores (RT Cores) and multiple encode/decode engines make it ideal for advanced visualization and digital twin applications.

The L40S GPU delivers up to 3.8x the real-time ray-tracing performance of its predecessor, and supports NVIDIA DLSS 3 for faster rendering and smoother frame rates. This makes the GPU ideal for developing applications on the NVIDIA Omniverse platform, enabling real-time, photorealistic 3D simulations and AI-enabled digital twins. With Omniverse on the L40S GPU, enterprises can develop advanced 3D applications and workflows for industrial digitalization that will allow them to design, simulate and optimize products, processes and facilities in real time before going into production.

OCI will offer the L40S GPU in its BM.GPU.L40S.4 bare-metal compute shape, featuring four NVIDIA L40S GPUs, each with 48GB of GDDR6 memory. This shape includes local NVMe drives with 7.38TB capacity, 4th Generation Intel Xeon CPUs with 112 cores and 1TB of system memory.

These shapes eliminate the overhead of any virtualization for high-throughput and latency-sensitive AI or machine learning workloads with OCI's bare-metal compute architecture. The accelerated compute shape features the NVIDIA BlueField-3 DPU for improved server efficiency, offloading data center tasks from CPUs to accelerate networking, storage and security workloads. The use of BlueField-3 DPUs furthers OCI's strategy of off-box virtualization across its entire fleet.

OCI Supercluster with NVIDIA L40S enables ultra-high performance with 800Gbps of internode bandwidth and low latency for up to 3,840 GPUs. OCI's cluster network uses NVIDIA ConnectX-7 NICs over RoCE v2 to support high-throughput and latency-sensitive workloads, including AI training.

We chose OCI AI infrastructure with bare-metal instances and NVIDIA L40S GPUs for 30% more efficient video encoding, said Sharon Carmel, CEO of Beamr Cloud. Videos processed with Beamr Cloud on OCI will have up to 50% reduced storage and network bandwidth consumption, speeding up file transfers by 2x and increasing productivity for end users. Beamr will provide OCI customers video AI workflows, preparing them for the future of video.

Single-GPU H100 VMs Coming Soon on OCI The VM.GPU.H100.1 compute virtual machine shape, accelerated by a single NVIDIA H100 Tensor Core GPU, is coming soon to OCI. This will provide cost-effective, on-demand access for enterprises looking to use the power of NVIDIA H100 GPUs for their generative AI and HPC workloads.

A single H100 provides a good platform for smaller workloads and LLM inference. For example, one H100 GPU can generate more than 27,000 tokens per second for Llama 3 8B (up to 4x more throughput than a single A100 GPU at FP16 precision) with NVIDIA TensorRT-LLM at an input and output sequence length of 128 and FP8 precision.

The VM.GPU.H100.1 shape includes 2 3.4TB of NVMe drive capacity, 13 cores of 4th Gen Intel Xeon processors and 246GB of system memory, making it well-suited for a range of AI tasks.

Oracle Cloud's bare-metal compute with NVIDIA H100 and A100 GPUs, low-latency Supercluster and high-performance storage delivers up to 20% better price-performance for Altair's computational fluid dynamics and structural mechanics solvers, said Yeshwant Mummaneni, chief engineer of data management analytics at Altair. We look forward to leveraging these GPUs with virtual machines for the Altair Unlimited virtual appliance.

GH200 Bare-Metal Instances Available for Validation OCI has also made available the BM.GPU.GH200 compute shape for customer testing. It features the NVIDIA Grace Hopper Superchip and NVLink-C2C, a high-bandwidth, cache-coherent 900GB/s connection between the NVIDIA Grace CPU and NVIDIA Hopper GPU. This provides over 600GB of accessible memory, enabling up to 10x higher performance for applications running terabytes of data compared to the NVIDIA A100 GPU.

Optimized Software for Enterprise AI Enterprises have a wide variety of NVIDIA GPUs to accelerate their AI, HPC and data analytics workloads on OCI. However, maximizing the full potential of these GPU-accelerated compute instances requires an optimized software layer.

NVIDIA NIM, part of the NVIDIA AI Enterprise software platform available on the OC
LINK: https://blogs.nvidia.com/blog/oracle-cloud-infrastructure-ai-gpu-digit...
See more stories from nvidia

Most recent headlines

05/01/2027

Worlds first 802.15.4ab-UWB chip verified by Calterah and Rohde & Schwarz to be demoed at CES 2026

Worlds first 802.15.4ab-UWB chip verified by Calterah and Rohde & Schwarz to be ...

04/08/2026

Dalet Announces Commercial Availability of Dalia, Bringing Media-Aware Agentic AI to Enterprise Productions

Dalet, a leading technology and service provider for media-rich organizations, t...

04/07/2026

Detective Conan: Fallen Angel of the Highway Opens in Dolby Cinemas Across Japan, Presented in Dolby Atmos and Dolby ...

April 7 2026, 19:00 (PDT) Detective Conan: Fallen Angel of the Highway Opens in...

01/06/2026

Dolby Sets the New Standard for Premium Entertainment at CES 2026

January 6 2026, 05:30 (PST) Dolby Sets the New Standard for Premium Entertainment at CES 2026 Throughout the week, Dolby brings to life the latest innovatio...

02/05/2026

Dalet Flex LTS Delivers Smarter Search, Faster Editing, and an AI-Ready Foundation for Modern Media

Dalet, a leading technology and service provider for media-rich organizations, t...

01/05/2026

NBCUniversal's Peacock to Be First Streamer to Integrate Dolby's Full Suite of Premium Picture and Sound Innovations

January 5 2026, 18:30 (PST) NBCUniversal's Peacock to Be First Streamer to ...

16/04/2026

NAB 2026: Appear Announces X5 General Availability, XM Management Tool, and SIx300 Module

Appear ASA (OSE: APR) will announce three additions to its X Platform at NAB Sho...

16/04/2026

CentralCast Deploys Harmonic XOS Media Processor for Public Broadcasting

Harmonic has announced that CentralCast, a centralized master control facility for U.S. public media, has deployed Harmonic's XOS Advanced Media Processor t...

16/04/2026

NAB 2026: Encompass Digital Media Integrates Interra Systems BATON into Cloud-Based Altitude Flow Platform

Interra Systems has announced that Encompass Digital Media has integrated its BA...

16/04/2026

NAB 2026: Grass Valley to Showcase Expanded Media Infrastructure Capabilities

Grass Valley will demonstrate its Media Infrastructure capabilities at NAB Show 2026 (Booth C2408, Central Hall), bringing routing, signal processing, and orche...

16/04/2026

Verizon Readies Robust Network Infrastructure Plans for FIFA World Cup 2026

As preparations ramp up for the FIFA World Cup 2026, Verizon has outlined a sweeping connectivity and infrastructure initiative that will underpin broadcast ope...

16/04/2026

Advanced Image Robotics Joins 2026 MLS Innovation Lab

Advanced Image Robotics (AIR) has announced its selection for the 2026 MLS Innovation Lab. AIR will work with MLS clubs, players, and executives on automated vi...

16/04/2026

NAB 2026: JWX Acquires True Anthem, Adding AI-Powered Social Publishing to Publisher Platform

JWX has announced the acquisition of True Anthem, an AI-powered social publishin...

16/04/2026

Jomboy Media and Fubo Launch 24/7 Creator-Led Channel

Jomboy Media and FuboTV Inc. have launched the Jomboy Media Channel, a 24/7 channel available to FuboTV base plan subscribers. The channel, timed to the start o...

16/04/2026

NAB 2026: Wave Central and EVS Announce SMPTE ST 2110 Interoperability Validation

Wave Central, a Domo Broadcast Company (Booth C2820), and EVS Broadcast Equipmen...

16/04/2026

Tata Communications and Formula 1 Release Race Before the Race, First Film in New Content Series

Tata Communications and Formula 1 have released Race Before the Race, the firs...

16/04/2026

NAB 2026: DPA Microphones N-Series Firmware Update Adds Duplex Gap and Guard Band Access for North American Users

DPA Microphones has released a firmware update for its N-Series Digital Wireless...

16/04/2026

NAB 2026: See Sony's Live Stage Presentations at NAB Show 2026

Sony's Live Stage at NAB Show 2026 is the place to hear directly from the content creators, end users, and technology experts who are pushing boundaries in ...

16/04/2026

Perfect Game and Youth Prospects Announce Broadcast Rights and Content Partnership

Perfect Game and Youth Prospects have announced a partnership covering broadcast...

16/04/2026

National Collegiate Rugby Partners with All Womens Sports Network for 2026 National 7s Championships

National Collegiate Rugby (NCR) has announced a media rights partnership with Al...

16/04/2026

NAB 2026: Audio-Technica Introduces BP350ST-UB and BP350ST-UL Mid-Side Stereo Broadcast Microphones

Audio-Technica has announced two new mid-side (MS) stereo broadcast microphones:...

16/04/2026

NAB 2026: PSSI Global Services Acquires Beagle Networks

PSSI Global Services has announced the acquisition of Beagle Networks, a provider of IT infrastructure and onsite technical support for media and enterprise cus...

16/04/2026

NAB 2026: Blackmagic Design Releases Blackmagic Camera for iOS 3.3

Blackmagic Design has released Blackmagic Camera for iOS 3.3, a free update available now from the Apple App Store. The update will be demonstrated at NAB Show ...

16/04/2026

SVG New Sponsor Spotlight: 4Wall Entertainment's Dave Caulwell on Scaling Live Event Production for Sports

As live sports production continues to expand across linear, digital, and in-ven...

16/04/2026

ESPN Enters First HDR Postseason With REMCO Support in New Bristol-Based Control Rooms

Part of an infrastructure upgrade, the recently constructed spaces accommodate n...

16/04/2026

NBC Sports Tips Off First NBA Playoffs Campaign in 23 Years With Six-Truck Fleet, Interchangeable Production Model

NBC has added NEP's Supershooter 11, which only just came online in time for...

16/04/2026

Introducing a Smarter, Smoother Experience for Tablets

At Spotify, we want your experience to feel intuitive and personal across every moment of the day. Whether you're streaming your favorite playlist while you...

16/04/2026

Telsie T - SonicWorlds extended classic German Equaliser

Vintage broadcast experts release second plug-in Telsie T is the second plug-in to be released by SonicWorld, a German audio company who specialise in servi...

16/04/2026

Softube unveil Flow Studio

Compact unit offers hands-on plug-in control Softube have just announced the launch of their latest hardware unit, the Flow Studio. Housed in a compact desk...

16/04/2026

Fiedler Audio release Armada

Use any VST3 plug-in on immersive audio Fiedler Audio have just released a powerful new plug-in wrapper that brings full VST3 processing to Dolby Atmos and ...

16/04/2026

Nugen Audio update Halo Vision

Analysis plug-in gains enhanced frequency readouts Nugen Audio's real-time analysis plug-in has just received a significant update that introduces some ...

16/04/2026

GearExpo UK Announced

Recording & Music Technology Show Sound On Sound are pleased to announce a new recording and music technology exhibition taking place in London on Saturday ...

16/04/2026

SBS reveals all-star alumni team to celebrate 70 years of the Eurovision Song Contest

SBS reveals all-star alumni team to celebrate 70 years of the Eurovision Song Co...

16/04/2026

Reimagining Earth Measurement for the AI Era: L3Harris and Xoople Develop a New Spaceborne Capability

For decades, understanding the physical world from space has required a trade-of...

16/04/2026

Locality and Nielsen Announce Landmark Integration of Media Data Engine, Transforming Local TV Measurement

The integration accelerates demographic audience delivery across local markets a...

16/04/2026

Quickplay Deploys Gray Medias New Streaming Platform

Share Copy link Facebook X Linkedin Bluesky Email...

16/04/2026

Spectrum TV App Launches On Google TV and Other Android TV Devices

Share Copy link Facebook X Linkedin Bluesky Email...

16/04/2026

Telestream Enables Multi-Cloud Media Workflows with Oracl...

Telestream Cloud Services, including Vantage Cloud, UP, and SENTRY monitoring tools, are now optimized for OCI, powering flexible multi-cloud media orchestratio...

16/04/2026

Fox Corporation Names Amazon Web Services its Preferred A...

Fox Corporation (Nasdaq: FOXA, FOX; FOX or the Company ) today announced a strategic collaboration with Amazon Web Services (AWS), naming AWS as its preferre...

16/04/2026

Triveni Digital Introduces New StreamScope Analyzer to Br...

Triveni Digital, a trusted leader in ATSC 1.0 and 3.0 service delivery, data broadcasting, and quality assurance solutions, today announced an ISDB-Tb capabilit...

16/04/2026

Synamedia and SoFast team up to accelerate FAST and OTT

Synamedia and SoFast announce strategic go-to-market partnership to accelerate FAST, pay-TV, and VOD At the NAB Show 2026, Synamedia and SoFast are announcing ...

16/04/2026

Imagine Builds on a Decade of ST 2110 Leadership at 2026...

Showcases New, Open-Standard IP Solutions Across Its Portfolio, From Production to Playout Imagine Communications is marking a decade of leadership in ST 2110 ...

16/04/2026

Roku Marks 100M Milestone

Share Copy link Facebook X Linkedin Bluesky Email...

16/04/2026

Sportway and Broadcast Solutions acquire Studio Automated...

Sportway Media Group, a world leading AI-automated sports production company and Broadcast Solutions, a leading system integrator and provider of innovative sol...

16/04/2026

AJA to Acquire Video Encoding Software Company Comprimato

Share Copy link Facebook X Linkedin Bluesky Email...

16/04/2026

NAB Show 2026: Sony Announces New Cameras, Virtual Production Tools

Share Copy link Facebook X Linkedin Bluesky Email...

16/04/2026

Synamedia launches GO Shorts

Share Copy link Facebook X Linkedin Bluesky Email...

16/04/2026

Clear-Com Introduces Arcadia and Eclipse HX Updates

Share Copy link Facebook X Linkedin Bluesky Email...

16/04/2026

AJA Enters into Agreement to Acquire Video Encoding Software Company Comprimato

AJA Enters into Agreement to Acquire Video Encoding Software Company Comprimato Brie Clayton April 15, 2026 0 Comments Deal will expand AJA's video ...

16/04/2026

Deity Announces PR-4 Compact Field Recorder with Pre-Orders Launching April 14

Deity Announces PR-4 Compact Field Recorder with Pre-Orders Launching April 14 Brie Clayton April 15, 2026 0 Comments Deity Microphones today announce...