Sony Pixel Power calrec Sony

Ray Shines with NVIDIA AI: Anyscale Collaboration to Help Developers Build, Tune, Train and Scale Production LLMs

18/09/2023

Large language model development is about to reach supersonic speed thanks to a collaboration between NVIDIA and Anyscale.

At its annual Ray Summit developers conference, Anyscale - the company behind the fast growing open-source unified compute framework for scalable computing - announced today that it is bringing NVIDIA AI to Ray open source and the Anyscale Platform. It will also be integrated into Anyscale Endpoints, a new service announced today that makes it easy for application developers to cost-effectively embed LLMs in their applications using the most popular open source models.

These integrations can dramatically speed generative AI development and efficiency while boosting security for production AI, from proprietary LLMs to open models such as Code Llama, Falcon, Llama 2, SDXL and more.

Developers will have the flexibility to deploy open-source NVIDIA software with Ray or opt for NVIDIA AI Enterprise software running on the Anyscale Platform for a fully supported and secure production deployment.

Ray and the Anyscale Platform are widely used by developers building advanced LLMs for generative AI applications capable of powering intelligent chatbots, coding copilots and powerful search and summarization tools.

NVIDIA and Anyscale Deliver Speed, Savings and Efficiency Generative AI applications are captivating the attention of businesses around the globe. Fine-tuning, augmenting and running LLMs requires significant investment and expertise. Together, NVIDIA and Anyscale can help reduce costs and complexity for generative AI development and deployment with a number of application integrations.

NVIDIA TensorRT-LLM, new open-source software announced last week, will support Anyscale offerings to supercharge LLM performance and efficiency to deliver cost savings. Also supported in the NVIDIA AI Enterprise software platform, Tensor-RT LLM automatically scales inference to run models in parallel over multiple GPUs, which can provide up to 8x higher performance when running on NVIDIA H100 Tensor Core GPUs, compared to prior-generation GPUs.

TensorRT-LLM automatically scales inference to run models in parallel over multiple GPUs and includes custom GPU kernels and optimizations for a wide range of popular LLM models. It also implements the new FP8 numerical format available in the NVIDIA H100 Tensor Core GPU Transformer Engine and offers an easy-to-use and customizable Python interface.

NVIDIA Triton Inference Server software supports inference across cloud, data center, edge and embedded devices on GPUs, CPUs and other processors. Its integration can enable Ray developers to boost efficiency when deploying AI models from multiple deep learning and machine learning frameworks, including TensorRT, TensorFlow, PyTorch, ONNX, OpenVINO, Python, RAPIDS XGBoost and more.

With the NVIDIA NeMo framework, Ray users will be able to easily fine-tune and customize LLMs with business data, paving the way for LLMs that understand the unique offerings of individual businesses.

NeMo is an end-to-end, cloud-native framework to build, customize and deploy generative AI models anywhere. It features training and inferencing frameworks, guardrailing toolkits, data curation tools and pretrained models, offering enterprises an easy, cost-effective and fast way to adopt generative AI.

Options for Open-Source or Fully Supported Production AI Ray open source and the Anyscale Platform enable developers to effortlessly move from open source to deploying production AI at scale in the cloud.

The Anyscale Platform provides fully managed, enterprise-ready unified computing that makes it easy to build, deploy and manage scalable AI and Python applications using Ray, helping customers bring AI products to market faster at significantly lower cost.

Whether developers use Ray open source or the supported Anyscale Platform, Anyscale's core functionality helps them easily orchestrate LLM workloads. The NVIDIA AI integration can help developers build, train, tune and scale AI with even greater efficiency.

Ray and the Anyscale Platform run on accelerated computing from leading clouds, with the option to run on hybrid or multi-cloud computing. This helps developers easily scale up as they need more computing to power a successful LLM deployment.

The collaboration will also enable developers to begin building models on their workstations through NVIDIA AI Workbench and scale them easily across hybrid or multi-cloud accelerated computing once it's time to move to production.

NVIDIA AI integrations with Anyscale are in development and expected to be available by the end of the year.

Developers can sign up to get the latest news on this integration as well as a free 90-day evaluation of NVIDIA AI Enterprise.

To learn more, attend the Ray Summit in San Francisco this week or watch the demo video below.

See this notice regarding NVIDIA's software roadmap.
LINK: https://blogs.nvidia.com/blog/2023/09/18/llm-anyscale-nvaie/...
See more stories from nvidia

Most recent headlines

01/04/2026

DOLBY AND DOUYIN EMPOWER THE NEXT GENERATON OF CREATORS WITH DOLBY VISION

January 4 2026, 18:00 (PST) DOLBY AND DOUYIN EMPOWER THE NEXT GENERATON OF CREATORS WITH DOLBY VISION Douyin Users Can Now Create And Share Videos With Stun...

03/01/2026

Release Rundown: What to Watch in January, From All That's Left of You to OBEX

A still from OBEX by Albert Birney, an official selection of the 2025 Sundance...

02/01/2026

Freezing Florida: NHL's EVP, Entertainment Bob Chesterman on Taking the Winter Classic to Miami

Freezing Florida: NHL's EVP, Entertainment Bob Chesterman on Taking the Wint...

02/01/2026

NHL Winter Classic 2026: TNT Sports Prepares for First NHL Outdoor Game in Sunshine State

NHL Winter Classic 2026: TNT Sports Prepares for First NHL Outdoor Game in Sunsh...

02/01/2026

ATSC To Showcase Latest NextGen TV Developments At CES 2026

Share Share by: Copy link Facebook X Whatsapp Pinterest Flipboard...

02/01/2026

Pearl TV To Unveil NextGen TV Converter Box Program At CES 2026

Share Share by: Copy link Facebook X Whatsapp Pinterest Flipboard...

02/01/2026

New RT documentary series takes us inside one of Ireland's busiest hospitals

Any Given Day: Cork University Hospital premieres Wednesday 7 January on RT One and RT Player at 9:35pm RT will debut a powerful new six-part documentary se...

02/01/2026

All episodes of Heated Rivalry will be landing on Sky and streaming service NOW on 10 January

Friday 2 January 2026 All episodes of Heated Rivalry will be landing on Sky and...

02/01/2026

Coming Soon to RT this New Year

Sequins, chat shows, live sporting action, ground-breaking docuseries and brand-new Irish drama to kick off 2026 New Year, New Content Coming Soon across RT ...

01/01/2026

Latin Grammy Cultural Foundation Announces 2026 Noel Schajris Scholarship

Latin Grammy Cultural Foundation Announces 2026 Noel Schajris Scholarship The scholarship will cover tuition and housing for one Berklee College of Music stud...

01/01/2026

GeForce NOW Rings In 2026 With 14 New Games in January

New year, new games, all with RTX 5080-powered cloud energy. GeForce NOW is kicking off 2026 by looking back at an unforgettable year of wins and wildly high fr...

31/12/2025

NFL Christmas Gameday on Netflix Scores Again With the Lions-Vikings Becoming the Most-Streamed NFL Game in US History

Back to All News NFL Christmas Gameday on Netflix Scores Again With the Lions-V...

30/12/2025

As the College Football Playoff Enters the Quarterfinals, ESPN Blows Out Its MegaCast Multiplatform Playbook

As the College Football Playoff Enters the Quarterfinals, ESPN Blows Out Its Meg...

30/12/2025

SVG's Best of 2025: Original Articles

SVG's Best of 2025: Original ArticlesTake a look back at all our coverage of big-time productions, game-changing technologies, and state-of-the-art new faci...

30/12/2025

L3Harris Sets Date for Fourth Quarter 2025 Earnings Release

MELBOURNE, Fla., Dec. 30, 2025 - L3Harris Technologies (NYSE: LHX) will release its fourth quarter 2025 financial results before the market opens on Thursday, J...

30/12/2025

TV Techs Most Popular Stories of 2025

Share Share by: Copy link Facebook X Whatsapp Pinterest Flipboard...

30/12/2025

Ring in the New Year with Patrick Kielty and a sparkling line-up on The Late Late Show New Year's Eve Show

Your live countdown to 2026 with Inhaler, David Gray, Lyra, Garron Noone, Sharon...

30/12/2025

Space42 Conducts Europe's First Licensed HAPS Flight

It marked the first civilian operational authorization for a HAPS flight in Europe, led by Space42's subsidiary, Mira Aerospace The flight demonstrated HAP...

29/12/2025

San Francisco 49ers Strike Gold With Halftime Laser Spectacular

San Francisco 49ers Strike Gold With Halftime Laser SpectacularStunning display caps $200 million renovation of Levi's Stadium techBy Dan Daley, Audio Edito...

29/12/2025

The Cup's Around the Corner: An Inside Look at Broadcast Preparations for the 2026 FIFA World Cup With FIFA's Oscar Sanchez

The Cup's Around the Corner: An Inside Look at Broadcast Preparations for th...

29/12/2025

SVG's Best of 2025: Longform Video

SVG's Best of 2025: Longform VideoWatch the standout keynote conversations, deep dives, and panel discussions from the year for free on SVG PLAY!By Brandon ...

29/12/2025

25 Ways Spotify Leveled Up Your Listening in 2025

From crisper Lossless audio and immersive music videos in beta to new Audiobooks+ plans, custom transitions between tracks, and in-app Messages, we keep levelin...

29/12/2025

TV Tech's Top Regulatory Stories of the Year

Share Share by: Copy link Facebook X Whatsapp Pinterest Flipboard...

27/12/2025

TV Tech's Top Data Dumps of 2025

Share Share by: Copy link Facebook X Whatsapp Pinterest Flipboard...

26/12/2025

TV Tech's Top Streaming Stories of 2025

Share Share by: Copy link Facebook X Whatsapp Pinterest Flipboard...

26/12/2025

TV Tech's Top Data Points of 2025

Share Share by: Copy link Facebook X Whatsapp Pinterest Flipboard...

25/12/2025

Make Spirits Bright With Holiday Hits on GeForce NOW

Holiday lights are twinkling, hot cocoa's on the stove and gamers are settling in for a well-earned break. Whether staying in or heading on a winter getawa...

24/12/2025

What is AI good for?

What is AI good for? Posted by MTI Film on December 24, 2025 What is AI good for? What is AI good for? It's been three years since ChatGPT first cap...

24/12/2025

AI in 2026: More Collaboration, Less Hype

Share Share by: Copy link Facebook X Whatsapp Pinterest Flipboard...

24/12/2025

Carr Lays Out FCCs 'Key Wins in 2025'

Share Share by: Copy link Facebook X Whatsapp Pinterest Flipboard...

24/12/2025

CES: Cineverse Unveils New Features for Cinesearch

Share Share by: Copy link Facebook X Whatsapp Pinterest Flipboard...

24/12/2025

IES, AES Promote Graham Kirk, Brienne Willcock

Share Share by: Copy link Facebook X Whatsapp Pinterest Flipboard...

24/12/2025

Ad Tech and CTV Experts Forecast 2026's Biggest Trends

Share Share by: Copy link Facebook X Whatsapp Pinterest Flipboard...

24/12/2025

Shaking up Sunday nights this January Dancing with the Stars unveils the new promo for 2026

RT has unveiled an exclusive first look at the new Dancing with the Stars promo...

24/12/2025

The Boyfriend' Season 2 Unveils Heartwarming Trailer, Key Art, and Participants' Profiles

Back to All News The Boyfriend' Season 2 Unveils Heartwarming Trailer, Key...

24/12/2025

Love, Fights, and Everything in Between: Badly in Love' Returns for Season 2

Back to All News Love, Fights, and Everything in Between: Badly in Love' Returns for Season 2 Entertainment 24 December 2025 GlobalJapan Link copied t...

24/12/2025

December 23, 2025

Scripps Research study links sleep variability with sleep apnea and hypertension How consumers' digital activity trackers could enable personalized health s...

23/12/2025

How guilas Cibaeas Dominican Winter League Games Are Locally Produced for Global Audience

How guilas Cibae as Dominican Winter League Games Are Locally Produced for Glob...

23/12/2025

CAMB.AI Enables European Athletics to Offer Multi-Language Support

CAMB.AI Enables European Athletics to Offer Multi-Language SupportPlan is to eventually offer translation into all languages spoken in EuropeBy Ken Kerschbaumer...

23/12/2025

Analysis: As Sports Media Values Trend Negative, Scarcity and Quality Are King

Analysis: As sports media values trend negative, scarcity and quality are king By Callum McCarthy, Editor-at-Large Monday, December 22, 2025 - 14:08 Print ...

23/12/2025

ESPN, Disney, and NBA Return to the Animated Altcast Fray With Second Edition of Dunk the Halls'

ESPN, Disney, and NBA Return to the Animated Altcast Fray With Second Edition of...

23/12/2025

End the Year on a High Note and Donate to the Sports Broadcasting Fund Today!

End the Year on a High Note and Donate to the Sports Broadcasting Fund Today!By Ken Kerschbaumer, Editorial Director Tuesday, December 23, 2025 - 12:25 pm P...

23/12/2025

Find Your Perfect Holiday Romance Listen With These Swoon-Worthy Audiobooks

The year is winding down, the weather outside is frightful, and it's the perfect time to escape into a story that warms the heart. For listeners looking for...

23/12/2025

L3Harris Receives Letter of Intent from Kratos Defense for Production of Large Hypersonic Solid Rocket Motors

A Zeus motor is hot fire tested at L3Harris' Camden, Arkansas, solid rocket ...

23/12/2025

FCC Bans All New Foreign-Made Drones

Share Share by: Copy link Facebook X Whatsapp Pinterest Flipboard...

23/12/2025

Gray Media Renews Its NBC Affiliation Agreements

Share Share by: Copy link Facebook X Whatsapp Pinterest Flipboard...

23/12/2025

Lightware to showcase breakthrough Google Meet and TPN MM...

Lightware will exhibit several major product innovations at ISE 2026, including the new USB-C BOOSTER-V1, Google Meet. integration for various Taurus UCX models...

23/12/2025

Nielsen, Roku Expand Measurement Partnership

Share Share by: Copy link Facebook X Whatsapp Pinterest Flipboard...