Sony Pixel Power calrec Sony

Oracle Cloud Infrastructure Expands NVIDIA GPU-Accelerated Instances for AI, Digital Twins and More

31/07/2024

Enterprises are rapidly adopting generative AI, large language models (LLMs), advanced graphics and digital twins to increase operational efficiencies, reduce costs and drive innovation.

However, to adopt these technologies effectively, enterprises need access to state-of-the-art, full-stack accelerated computing platforms. To meet this demand, Oracle Cloud Infrastructure (OCI) today announced NVIDIA L40S GPU bare-metal instances available to order and the upcoming availability of a new virtual machine accelerated by a single NVIDIA H100 Tensor Core GPU. This new VM expands OCI's existing H100 portfolio, which includes an NVIDIA HGX H100 8-GPU bare-metal instance.

Paired with NVIDIA networking and running the NVIDIA software stack, these platforms deliver powerful performance and efficiency, enabling enterprises to advance generative AI.

NVIDIA L40S Now Available to Order on OCI The NVIDIA L40S is a universal data center GPU designed to deliver breakthrough multi-workload acceleration for generative AI, graphics and video applications. Equipped with fourth-generation Tensor Cores and support for the FP8 data format, the L40S GPU excels in training and fine-tuning small- to mid-size LLMs and in inference across a wide range of generative AI use cases.

For example, a single L40S GPU (FP8) can generate up to 1.4x more tokens per second than a single NVIDIA A100 Tensor Core GPU (FP16) for Llama 3 8B with NVIDIA TensorRT-LLM at an input and output sequence length of 128.

The L40S GPU also has best-in-class graphics and media acceleration. Its third-generation NVIDIA Ray Tracing Cores (RT Cores) and multiple encode/decode engines make it ideal for advanced visualization and digital twin applications.

The L40S GPU delivers up to 3.8x the real-time ray-tracing performance of its predecessor, and supports NVIDIA DLSS 3 for faster rendering and smoother frame rates. This makes the GPU ideal for developing applications on the NVIDIA Omniverse platform, enabling real-time, photorealistic 3D simulations and AI-enabled digital twins. With Omniverse on the L40S GPU, enterprises can develop advanced 3D applications and workflows for industrial digitalization that will allow them to design, simulate and optimize products, processes and facilities in real time before going into production.

OCI will offer the L40S GPU in its BM.GPU.L40S.4 bare-metal compute shape, featuring four NVIDIA L40S GPUs, each with 48GB of GDDR6 memory. This shape includes local NVMe drives with 7.38TB capacity, 4th Generation Intel Xeon CPUs with 112 cores and 1TB of system memory.

These shapes eliminate the overhead of any virtualization for high-throughput and latency-sensitive AI or machine learning workloads with OCI's bare-metal compute architecture. The accelerated compute shape features the NVIDIA BlueField-3 DPU for improved server efficiency, offloading data center tasks from CPUs to accelerate networking, storage and security workloads. The use of BlueField-3 DPUs furthers OCI's strategy of off-box virtualization across its entire fleet.

OCI Supercluster with NVIDIA L40S enables ultra-high performance with 800Gbps of internode bandwidth and low latency for up to 3,840 GPUs. OCI's cluster network uses NVIDIA ConnectX-7 NICs over RoCE v2 to support high-throughput and latency-sensitive workloads, including AI training.

We chose OCI AI infrastructure with bare-metal instances and NVIDIA L40S GPUs for 30% more efficient video encoding, said Sharon Carmel, CEO of Beamr Cloud. Videos processed with Beamr Cloud on OCI will have up to 50% reduced storage and network bandwidth consumption, speeding up file transfers by 2x and increasing productivity for end users. Beamr will provide OCI customers video AI workflows, preparing them for the future of video.

Single-GPU H100 VMs Coming Soon on OCI The VM.GPU.H100.1 compute virtual machine shape, accelerated by a single NVIDIA H100 Tensor Core GPU, is coming soon to OCI. This will provide cost-effective, on-demand access for enterprises looking to use the power of NVIDIA H100 GPUs for their generative AI and HPC workloads.

A single H100 provides a good platform for smaller workloads and LLM inference. For example, one H100 GPU can generate more than 27,000 tokens per second for Llama 3 8B (up to 4x more throughput than a single A100 GPU at FP16 precision) with NVIDIA TensorRT-LLM at an input and output sequence length of 128 and FP8 precision.

The VM.GPU.H100.1 shape includes 2 3.4TB of NVMe drive capacity, 13 cores of 4th Gen Intel Xeon processors and 246GB of system memory, making it well-suited for a range of AI tasks.

Oracle Cloud's bare-metal compute with NVIDIA H100 and A100 GPUs, low-latency Supercluster and high-performance storage delivers up to 20% better price-performance for Altair's computational fluid dynamics and structural mechanics solvers, said Yeshwant Mummaneni, chief engineer of data management analytics at Altair. We look forward to leveraging these GPUs with virtual machines for the Altair Unlimited virtual appliance.

GH200 Bare-Metal Instances Available for Validation OCI has also made available the BM.GPU.GH200 compute shape for customer testing. It features the NVIDIA Grace Hopper Superchip and NVLink-C2C, a high-bandwidth, cache-coherent 900GB/s connection between the NVIDIA Grace CPU and NVIDIA Hopper GPU. This provides over 600GB of accessible memory, enabling up to 10x higher performance for applications running terabytes of data compared to the NVIDIA A100 GPU.

Optimized Software for Enterprise AI Enterprises have a wide variety of NVIDIA GPUs to accelerate their AI, HPC and data analytics workloads on OCI. However, maximizing the full potential of these GPU-accelerated compute instances requires an optimized software layer.

NVIDIA NIM, part of the NVIDIA AI Enterprise software platform available on the OC
LINK: https://blogs.nvidia.com/blog/oracle-cloud-infrastructure-ai-gpu-digit...
See more stories from nvidia

Most recent headlines

05/01/2027

Worlds first 802.15.4ab-UWB chip verified by Calterah and Rohde & Schwarz to be demoed at CES 2026

Worlds first 802.15.4ab-UWB chip verified by Calterah and Rohde & Schwarz to be ...

04/08/2026

Dalet Announces Commercial Availability of Dalia, Bringing Media-Aware Agentic AI to Enterprise Productions

Dalet, a leading technology and service provider for media-rich organizations, t...

04/07/2026

Detective Conan: Fallen Angel of the Highway Opens in Dolby Cinemas Across Japan, Presented in Dolby Atmos and Dolby ...

April 7 2026, 19:00 (PDT) Detective Conan: Fallen Angel of the Highway Opens in...

01/06/2026

Dolby Sets the New Standard for Premium Entertainment at CES 2026

January 6 2026, 05:30 (PST) Dolby Sets the New Standard for Premium Entertainment at CES 2026 Throughout the week, Dolby brings to life the latest innovatio...

28/05/2026

Calrec Scales ImPulseV to Expand Choice in Virtualized Audio Workflows

Share Copy link Facebook X Linkedin Bluesky Email...

28/05/2026

Berliner Ensemble Upgrades Backstage Infrastructure With...

Riedel Communications today announced that Berliner Ensemble, one of Berlin's five major theater companies, has expanded its backstage communications and te...

28/05/2026

Digital Alert Systems Ed Czarnecki to Speak on Advanced E...

Company's VP Government and International to Address How Broadcasters Can Operationalize Advanced Emergency Information for NextGen TV Digital Alert Syst...

28/05/2026

Radio BGM Upgrades and Futureproofs with DHD SX2 Audio Mi...

Radio BGM (https://radiobgm.org.uk/), Llanellis multiple-award-winning community and hospital radio station, has invested in a DHD SX2 audio mixing console and ...

28/05/2026

Calrec Scales ImPulseV to Empower Broadcasters with Great...

Further strengthening its virtualisation strategy to fully support broadcasters as they enter a new broadcast age, Calrec announces an expansion to its ImPulseV...

28/05/2026

NAB, MPA and NCTA Defend Current TV Ratings System

Share Copy link Facebook X Linkedin Bluesky Email...

28/05/2026

Amazon MGM Studios, AWS Launch the GenAI Creators' Fund

Share Copy link Facebook X Linkedin Bluesky Email...

28/05/2026

AudioShake Launches New Copyright Compliance System

Share Copy link Facebook X Linkedin Bluesky Email...

28/05/2026

ITV Studios Deploys Cuez Live Production Platform

Share Copy link Facebook X Linkedin Bluesky Email...

28/05/2026

ACT Entertainment Introduces Green Hippo Estuary Series a...

Green Hippo, an ACT Entertainment brand, will pull back the curtain on the Estuary Series, a next-generation media control platform engineered for the scale, sp...

28/05/2026

ZEISS Introduces Panoptes 65 Cinema Lenses at Cine Gear E...

First Hands-On Opportunity at Universal Studios Lot, June 5-6 Oberkochen, Germany, May 27, 2026 ZEISS is launching their new Panoptes 65 primes at Cine Gear ...

28/05/2026

Wooden Camera Releases New Accessories for Blackmagic URS...

Irvine, CA May 20, 2026 Wooden Camera today announced the release of new accessories for the Blackmagic URSA Cine Immersive. The new lineup includes a redes...

28/05/2026

Six films, one unmissable collection. First Facts documentaries debut on 10 Streaming

Six films, one unmissable collection. First Facts documentaries debut on 10 Stre...

28/05/2026

Emma Watkins brings all-new Emma's Dance Club to ABC Kids

Emma Watkins brings all-new Emma's Dance Club to ABC Kids 28 May 2026 Emmas Dance Club. Image credit: Sarah Wilson. The ABC, Screen Australia and Screen N...

28/05/2026

Cool Concentric Text for Cavalry

Cool Concentric Text for Cavalry Simon Ubsdell May 27, 2026 0 Comments In this tutorial I'll show you how to make this popular text effect that is...

28/05/2026

The Name's Gaming Cloud Gaming: 007 First Light' Launches on GeForce NOW

License to stream, shaken and stirred. GeForce NOW is dialing up the espionage with the launch of 007 First Light, letting members slip into James Bond's r...

28/05/2026

NVIDIA Research Advances Robotics From Simulation to the Real World

Robotics is entering a new phase: moving from controlled demos and scripted automation toward generalizable, reliable embodied autonomy in the real world. At ...

28/05/2026

May 21, 2026

Scripps Research's Skaggs Graduate School awards doctoral degrees to 34th graduating class May 21, 2026 Scripps Research's Skaggs Graduate School of ...

28/05/2026

May 27, 2026

Scripps Research chemist Jin-Quan Yu is named a Fellow of the Royal Society Yu is honored by the U.K.'s national academy of sciences for his work in synthet...

27/05/2026

Telestream Appoints Benjamin Desbois as CEO, Effective July 1

Telestream has announced that its Board of Directors has appointed Benjamin Desbois as Chief Executive Officer, effective July 1, 2026. Desbois, currently Teles...

27/05/2026

FOX MLB Leads Live-Event Categories; ESPN Is Tops Overall at 47th Annual Sports Emmy Awards

ESPN garnered 10 awards; NBC's Sunday Night Football received the Outstandin...

27/05/2026

Matrox Video Marks 50th Anniversary, Announces New Product Launch for June

Matrox Video is celebrating its 50th anniversary, marking five decades of operations from its headquarters in Montreal, Canada. Founded in 1976, the company has...

27/05/2026

MLB Announces Fan Engagement Initiatives for Americas 250th Anniversary

Major League Baseball has announced a series of initiatives tied to America's Semiquincentennial, including a national marketing campaign, Fourth of July br...

27/05/2026

Advanced Systems Group Hires Brian Gross as Account Manager for Audio Team

Advanced Systems Group (ASG) has announced that Brian Gross has joined the company as an Account Manager on its Audio team, based in the Burbank office. He will...

27/05/2026

Nielsen Research: Hispanic Fans, Asian Markets Drive Global Soccer Audience Ahead of World Cup 2026

Nielsen has released new research on soccer fandom ahead of the FIFA World Cup 2...

27/05/2026

ESL FACEIT Group Debuts First Ever Esports Vertical Stream Co-Developed With TikTok

ESL FACEIT Group (EFG) has unveiled a new partnership with TikTok to bring broad...

27/05/2026

Two Weeks Away: FIFA Outlines Production Plans for Highly Anticipated North American-Based World Cup

FIFA's Oscar Sanchez gives a deeper look to how this tournament will be cove...

27/05/2026

SVG Students To Watch: Maggie Lynn, Virginia Tech

The soon-to-be senior from Charlottesville is building her skills in replay, TD, and even creative content for HokieVision and its ACC Network productions In t...

27/05/2026

A Global Festival of Football: FOX Sports Illustrates Strategy to Bring Every FIFA Mens World Cup Match to the U.S. Audience

FOX Sports' Mike Davies breaks down the vision for this summer's showcas...

27/05/2026

Top-Tier Storytelling: Host Broadcast Services Works at Capturing the Atmosphere of the FIFA Mens World Cup

HBS's Paul King, FIFA's Oscar Sanchez preview how the masses at home wil...

27/05/2026

Matt Gangl & Pete Macheska on FOX MLBs Huge Night and an Unforgettable Postseason Run

FOX's MLB coverage dominated the night at the 47th Annual Sports Emmy Awards...

27/05/2026

FOXs Mike Davies and Team on Outstanding Technical Team Win for 2025 World Series

One of the most memorable Postseasons in baseball history would have had no memo...

27/05/2026

NBC Sports Rob Hyland Reflects on an Unforgettable Sunday Night Football Season

NBC's Sunday Night Football is among the most decorated and most watched programs in the history of television. It added to its jam-packed trophy case on Tu...

27/05/2026

Prime Videos John Ward and Mike Francis on Groundbreaking NBA on Prime Video Studio

The 2026 Sports Emmys marked a watershed moment for Prime Video Sports. After bu...

27/05/2026

Countdown to FIFA World Cup 2026: SVG Launches SportsTechLive Blog in Lead-up to Winter Games

With the Opening Match just over two weeks away, the entire sports-production-te...

27/05/2026

Spotify Brings Long-Form Magazine Articles to Audio

Spotify already brings together listeners' favorite music, podcasts, and audiobooks in one place. Now, we're trialing a new format that expands the cont...

27/05/2026

Podcast Clips Make Your Favorite Moments Easier to Save and Share

The best podcast moments deserve more than just a mental note. That's why today, we're making those moments easier to save and share with clips. Whethe...

27/05/2026

Spotify and Netflix Partner With Jay Shetty to Bring On Purpose' to Video Across Both Platforms

On Purpose is one of the most popular podcasts in the world, known for conversat...

27/05/2026

Olivia Rodrigo Brings Billions Club Live to Barcelona: Watch the Concert Film Now

On May 8, 1,500 of Olivia Rodrigo's top fans gathered in Barcelona's Tea...

27/05/2026

JZ Microphones announce the MU-1

Hybrid design combines large-diaphragm capsule & ribbon JZ Microphones have teamed up with Grammy-winning producer and engineer Marc Urselli to develop a ne...

27/05/2026

Tape Effects Collection from AIR Music Tech

Three new plug-ins inspired by classic tape effects AIR Music Tech's latest release delivers a set of plug-ins that aim to capture the character, moveme...

27/05/2026

The Crow Hill Company's Absurdly Quiet Piano goes Pro

Piano played on the edge of silence The Crow Hill Company's Vaults collection offers a continual rotation of instruments that are given away for free fo...

27/05/2026

Arturia release Memory V

Recreates Moog's iconic Memorymoog polysynth Arturia's vast software instrument range offers a combination of new and old, with innovative modern so...

27/05/2026

Accentize introduce free dxLevel plug-in

Offers loudness levelling for speech and dialogue Accentize have built up a solid reputation with their audio-restoration tools, and their latest plug-in is...

27/05/2026

10,000 units strong - The Rohde & Schwarz R&S M3SR Radio 4400

10,000 units strong - The Rohde & Schwarz R&S M3SR Radio 4400 Rohde & Schwarz celebrates a major manufacturing milestone, producing its 10,000th R&S M3SR Radi...