Sony Pixel Power calrec Sony

Ray Shines with NVIDIA AI: Anyscale Collaboration to Help Developers Build, Tune, Train and Scale Production LLMs

18/09/2023

Large language model development is about to reach supersonic speed thanks to a collaboration between NVIDIA and Anyscale.

At its annual Ray Summit developers conference, Anyscale - the company behind the fast growing open-source unified compute framework for scalable computing - announced today that it is bringing NVIDIA AI to Ray open source and the Anyscale Platform. It will also be integrated into Anyscale Endpoints, a new service announced today that makes it easy for application developers to cost-effectively embed LLMs in their applications using the most popular open source models.

These integrations can dramatically speed generative AI development and efficiency while boosting security for production AI, from proprietary LLMs to open models such as Code Llama, Falcon, Llama 2, SDXL and more.

Developers will have the flexibility to deploy open-source NVIDIA software with Ray or opt for NVIDIA AI Enterprise software running on the Anyscale Platform for a fully supported and secure production deployment.

Ray and the Anyscale Platform are widely used by developers building advanced LLMs for generative AI applications capable of powering intelligent chatbots, coding copilots and powerful search and summarization tools.

NVIDIA and Anyscale Deliver Speed, Savings and Efficiency Generative AI applications are captivating the attention of businesses around the globe. Fine-tuning, augmenting and running LLMs requires significant investment and expertise. Together, NVIDIA and Anyscale can help reduce costs and complexity for generative AI development and deployment with a number of application integrations.

NVIDIA TensorRT-LLM, new open-source software announced last week, will support Anyscale offerings to supercharge LLM performance and efficiency to deliver cost savings. Also supported in the NVIDIA AI Enterprise software platform, Tensor-RT LLM automatically scales inference to run models in parallel over multiple GPUs, which can provide up to 8x higher performance when running on NVIDIA H100 Tensor Core GPUs, compared to prior-generation GPUs.

TensorRT-LLM automatically scales inference to run models in parallel over multiple GPUs and includes custom GPU kernels and optimizations for a wide range of popular LLM models. It also implements the new FP8 numerical format available in the NVIDIA H100 Tensor Core GPU Transformer Engine and offers an easy-to-use and customizable Python interface.

NVIDIA Triton Inference Server software supports inference across cloud, data center, edge and embedded devices on GPUs, CPUs and other processors. Its integration can enable Ray developers to boost efficiency when deploying AI models from multiple deep learning and machine learning frameworks, including TensorRT, TensorFlow, PyTorch, ONNX, OpenVINO, Python, RAPIDS XGBoost and more.

With the NVIDIA NeMo framework, Ray users will be able to easily fine-tune and customize LLMs with business data, paving the way for LLMs that understand the unique offerings of individual businesses.

NeMo is an end-to-end, cloud-native framework to build, customize and deploy generative AI models anywhere. It features training and inferencing frameworks, guardrailing toolkits, data curation tools and pretrained models, offering enterprises an easy, cost-effective and fast way to adopt generative AI.

Options for Open-Source or Fully Supported Production AI Ray open source and the Anyscale Platform enable developers to effortlessly move from open source to deploying production AI at scale in the cloud.

The Anyscale Platform provides fully managed, enterprise-ready unified computing that makes it easy to build, deploy and manage scalable AI and Python applications using Ray, helping customers bring AI products to market faster at significantly lower cost.

Whether developers use Ray open source or the supported Anyscale Platform, Anyscale's core functionality helps them easily orchestrate LLM workloads. The NVIDIA AI integration can help developers build, train, tune and scale AI with even greater efficiency.

Ray and the Anyscale Platform run on accelerated computing from leading clouds, with the option to run on hybrid or multi-cloud computing. This helps developers easily scale up as they need more computing to power a successful LLM deployment.

The collaboration will also enable developers to begin building models on their workstations through NVIDIA AI Workbench and scale them easily across hybrid or multi-cloud accelerated computing once it's time to move to production.

NVIDIA AI integrations with Anyscale are in development and expected to be available by the end of the year.

Developers can sign up to get the latest news on this integration as well as a free 90-day evaluation of NVIDIA AI Enterprise.

To learn more, attend the Ray Summit in San Francisco this week or watch the demo video below.

See this notice regarding NVIDIA's software roadmap.
LINK: https://blogs.nvidia.com/blog/2023/09/18/llm-anyscale-nvaie/...
See more stories from nvidia

Most recent headlines

11/12/2025

Lawo, SMPTE To Conduct ST 2110 Practical Lab

RASTATT, Germany Lawo and the Society of Motion Picture and Television Engineers (SMPTE) have partnered to launch the SMPTE ST 2110 Practical Lab, an immersive ...

11/12/2025

Comcast's Xfinity Revamps National Video Plans

PHILADELPHIA Comcasts Xfinity operating brand has announced the launch of new national video plans with all-in pricing that the operator said will provide custo...

11/12/2025

Analyst: Pay TV Video Subs Rise for First Time Since 2017

After eight years of declines, MoffettNathansons new Cord Cutting Monitor for Q3 2025 shows that pay TV subscribers to linear TV packages rose by 303,000, the f...

11/12/2025

Happy Holidays from Berklee

Happy Holidays from Berklee Enjoy this years holiday student-performance video. December 10, 2025 By Office of the President Dear Berklee community, As w...

10/12/2025

Sound-Alike Commercials Are Part of Sports' Soundtrack

Sound-Alike Commercials Are Part of Sports' Soundtrack Johnny Cash for Coca-Cola is the latest in a long litany of sonic approximationsBy Dan Daley, Audio ...

10/12/2025

Immersive Sound Is Logical Next Step for Sports Venues

Immersive Sound Is Logical Next Step for Sports VenuesSound-systems suppliers are sanguine, but the market has its challengesBy Dan Daley, Audio Editor Wednes...

10/12/2025

The Romans Built Arenas for Immersive Sound 2,000 Years Ago

The Romans Built Arenas for Immersive Sound 2,000 Years AgoThe historic Arena of Nimes in France is still in use todayBy Dan Daley, Audio Editor Wednesday, De...

10/12/2025

SVG Summit 2025 Preview: Audio Workshop Hits on Immersive, Virtualized, and Next-Gen Streaming Workflows

SVG Summit 2025 Preview: Audio Workshop Hits on Immersive, Virtualized, and Next...

10/12/2025

SVG Summit 2025 Technology Exhibits Preview: Audio Spotlight

SVG Summit 2025 Technology Exhibits Preview: Audio SpotlightBy SVG Staff Wednesday, December 10, 2025 - 8:21 am Print This Story | Subscribe Story Highlig...

10/12/2025

SVG Europe Audio: Listening to the Sounds of Powder and Ice at Milano Cortina with a Behind the Scenes Tour of OBS and NBC's Audio Set Ups

SVG Europe Audio: Listening to the sounds of powder and ice at Milano Cortina wi...

10/12/2025

Advancements in Audio Technology: Capturing the Atmosphere of Live Sports

Advancements in audio technology: Capturing the atmosphere of live sports By David Davies Tuesday, November 25, 2025 - 09:27 Print This Story Although wor...

10/12/2025

Everything Smelled of Popcorn: The Art of Bringing the Complex Sound of Esports to Fans With Sound Supervisor Matt Gilbert

Everything smelled of popcorn: The art of bringing the complex sound of esports ...

10/12/2025

2026 Sundance Film Festival Unveils 97 Projects Selected for the Feature Film and Episodic Program

Top L-R: Ha-Chan, Shake Your Booty!, Hanging by a Wire, Broken English, Buddy C...

10/12/2025

You're in Control: Spotify Lets You Steer the Algorithm

For the first time, Spotify is giving users the power to steer the algorithm. Gustav S derstr m, Spotify's Co-President, CPO, and CTO, shares the vision beh...

10/12/2025

L3Harris to Produce Additional Solid Rocket Motors for Precision-Guided Artillery System

L3Harris' new contract for Guided Multiple Launch Rocket System Insensitive ...

10/12/2025

US Space Force Expands Offensive Space Programs Through L3Harris Foreign Sales

L3Harris Meadowlands system has been designed with an open architecture software system that allows for more flexible and efficient software updates. This capab...

10/12/2025

Football Shifts TV Viewing Towards Ad Supported, Nielsen's Q3 2025 Ad Supported Gauge Finds

During this interval, streaming comprised the majority of ad supported TV (46.4%...

10/12/2025

Bitcentral Names Venture Capital Exec Rick Arnold to Board

NEWPORT BEACH, Calif. Bitcentral, a provider of production, asset management, playout and streaming workflow solutions, has named technology veteran Rick Arnold...

10/12/2025

TV Tech Announces Winners of 2025 Best in Market Awards for M&E Tech

TV Tech is delighted to reveal the winners of the 2025 Media & Entertainment: Best in Market Awards....

10/12/2025

AIMS, VSF, AMWA, EBU To Hold Inaugural IPMX Testing, Certification Event

BOTHELL, Wash. The Alliance for IP Media Solutions (AIMS), the Video Services Forum (VSF), the Advanced Media Workflow Association (AMWA) and the European Broad...

10/12/2025

DirecTV Launches Peacock Games

In a notable example of how pay TV operators are integrating streaming services into their lineup and using those services to retain or attract subscribers, Dir...

10/12/2025

Chaos Brings Real-Time Rendering to Maya and Houdini

Today, Chaos builds instant feedback into the viewport, connecting Maya and Houdini to Chaos Vantage's real-time path tracer. Artists can now assess 3D asse...

10/12/2025

Smeup doubles capacity with Cubbit under a new agreement...

Smeup, a key partner for companies engaged in digital transformation, today announced the expansion of its adoption of Cubbit, the first geo-distributed cloud s...

10/12/2025

Mediagenix Strengthens Its Security Posture with ISO 2700...

Mediagenix, a global leader in smart content solutions to profitably connect the right content to the right audience, today announced two significant milestones...

10/12/2025

HDR10+ Technologies Unveils HDR10+ ADVANCED Dynamic Metadata Technology

BEAVERTON, Ore. HDR10+ Technologies, LLC has announced that they will soon begin the licensing and certification of devices, content, and services that support ...

10/12/2025

SMPTE, EBU, ETC Publish Report on AI's Impact on the Media

SMPTE has joined forces with the European Broadcasting Union (EBU) and Entertainment Technology Center (ETC) to publish an updated report on AI and its impact o...

10/12/2025

Clear-Com Appoints Kris Koch as New Director of Sales - N...

Clear-Com is pleased to announce the appointment of Kris Koch as Director of Sales - North & South America. In this expanded leadership role, Kris will oversee...

10/12/2025

Mavis Camera Launches Film Kit Unlocking LUT Workflows an...

Mavis today announced the latest version of Mavis Camera (v7.4), a major update to its professional iOS camera app, headlined by the launch of Film Kit - an opt...

10/12/2025

Creamsource Taps Industry Heavyweight Markus Zeiler as Gr...

Creamsource, renowned for its Vortex series of cinematic lighting, is laying the groundwork for its next phase of growth with the addition of Markus Zeiler as G...

10/12/2025

Digital Alert Systems Introduces DAS3-DC-PS DASDEC-III DC...

Digital Alert Systems, a global leader in emergency communications solutions for media providers, today announced that the DAS3-DC-PS, a new DC power supply opt...

10/12/2025

Riedel and Racing Electronics Announce Strategic Partners...

Riedel Communications today announced it has formed a strategic partnership with Racing Electronics, a premier provider of motorsport communication equipment in...

10/12/2025

GALSNGEAR Announces 2026 Leadership Retreats on East and...

#GALSNGEAR is launching two signature leadership retreats in early 2026, designed to equip women in media, entertainment, and technology with the tools to lead...

10/12/2025

CVP Launches Global Price Guarantee for Seamless Internat...

Providing worldwide customers with total confidence through transparent, all-inclusive pricing CVP, one of Europe's leading suppliers of professional video...

10/12/2025

Securing the Future of Broadcast TV in the U.S.

With the Federal Communications Commission working on new rules for the deployment of NextGen TV, next year promises to be an important one for both the future ...

10/12/2025

Former Charter CEO Tom Rutledge to Receive Cable Centers Bresnan Award

DENVER Tom Rutledge, director emeritus and former president and CEO of Charter Communications, will be honored with the 2026 Bresnan Ethics in Business Award by...

10/12/2025

Cadent Acquires YouTube Measurement Firm VuePlanner

NEW YORK Novocap's Cadent has acquired VuePlanner, a YouTube video ad planning, optimization, and measurement company in a deal that will help Cadent expand...

10/12/2025

3 Ways NVIDIA Is Powering the Industrial Revolution

The NVIDIA accelerated computing platform is leading supercomputing benchmarks once dominated by CPUs, enabling AI, science, business and computing efficiency w...

10/12/2025

How NVIDIA H100 GPUs on CoreWeave's AI Cloud Platform Delivered a Record-Breaking Graph500 Run

The world's top-performing system for graph processing at scale was built on...

10/12/2025

Opt-In NVIDIA Software Enables Data Center Fleet Management

As the scale and complexity of AI infrastructure grows, data center operators need continuous visibility into factors including performance, temperature and pow...

10/12/2025

Avoid Playlist Conflicts: Scheduling Back-to-Back Special Playlists

In preparation for the madness of March, here are some important reminders for scheduling back-to-back Special Playlists. The first Special Playlist MUST end b...

10/12/2025

VEON's Rising Capital Markets Profile Strengthened by Inclusion in Key Global Indices

10 Dec 2025 VEON's Rising Capital Markets Profile Strengthened by Inclusion...

10/12/2025

VEON Recognized for JazzCash, Kyivstar and Jazz at the World Communication Awards 2025

10 Dec 2025 VEON Recognized for JazzCash, Kyivstar and Jazz at the World Commun...

10/12/2025

Tribeca Films to Release the Independent Documentary Film Beam Me Up, Sulu by Timour Gregory and Sasha Schneider

December 10th, 2025 TRIBECA FILMS TO RELEASE THE INDEPENDENT DOCUMENTARY FILM...

10/12/2025

Sky extends partnership with the Ladies European Tour for a landmark 30th year

Wednesday 10 December 2025 Sky extends partnership with the Ladies European Tour for a landmark 30th year Sky and the Ladies European Tour (LET) have announce...

10/12/2025

Walk-on if you love the darts: James Maddison, Luke Littler and Big John star as Club 180 opens before 2026 PDC World Darts Championship

Wednesday 10 December 2025 Walk-on if you love the darts: James Maddison, Luke ...

10/12/2025

Rohde & Schwarz presents world's first RF power sensor with 0.80 mm RF connector for gapless DC to 150 GHz coverage

Rohde & Schwarz presents world's first RF power sensor with 0.80 mm RF conne...

10/12/2025

2026 Starts With a Swoon: Kim Seon-ho and Go Youn-jung Lead Can This Love Be Translated?', Premiering January 16

Back to All News 2026 Starts With a Swoon: Kim Seon-ho and Go Youn-jung Lead C...

10/12/2025

'Berlin and the Lady with an Ermine' Arrives to Netflix on May 15

Back to All News Berlin and the Lady with an Ermine Arrives to Netflix on May 15 Entertainment 10 December 2025 GlobalSpain Link copied to clipboard THE N...

10/12/2025

From stand-up to Foxtrot: Comedian Michael Fry revealed as fourth contestant for Dancing with the Stars 2026

It's out of the frying pan and into the sequins for comedian and actor Micha...