
Large language model development is about to reach supersonic speed thanks to a collaboration between NVIDIA and Anyscale.
At its annual Ray Summit developers conference, Anyscale - the company behind the fast growing open-source unified compute framework for scalable computing - announced today that it is bringing NVIDIA AI to Ray open source and the Anyscale Platform. It will also be integrated into Anyscale Endpoints, a new service announced today that makes it easy for application developers to cost-effectively embed LLMs in their applications using the most popular open source models.
These integrations can dramatically speed generative AI development and efficiency while boosting security for production AI, from proprietary LLMs to open models such as Code Llama, Falcon, Llama 2, SDXL and more.
Developers will have the flexibility to deploy open-source NVIDIA software with Ray or opt for NVIDIA AI Enterprise software running on the Anyscale Platform for a fully supported and secure production deployment.
Ray and the Anyscale Platform are widely used by developers building advanced LLMs for generative AI applications capable of powering intelligent chatbots, coding copilots and powerful search and summarization tools.
NVIDIA and Anyscale Deliver Speed, Savings and Efficiency Generative AI applications are captivating the attention of businesses around the globe. Fine-tuning, augmenting and running LLMs requires significant investment and expertise. Together, NVIDIA and Anyscale can help reduce costs and complexity for generative AI development and deployment with a number of application integrations.
NVIDIA TensorRT-LLM, new open-source software announced last week, will support Anyscale offerings to supercharge LLM performance and efficiency to deliver cost savings. Also supported in the NVIDIA AI Enterprise software platform, Tensor-RT LLM automatically scales inference to run models in parallel over multiple GPUs, which can provide up to 8x higher performance when running on NVIDIA H100 Tensor Core GPUs, compared to prior-generation GPUs.
TensorRT-LLM automatically scales inference to run models in parallel over multiple GPUs and includes custom GPU kernels and optimizations for a wide range of popular LLM models. It also implements the new FP8 numerical format available in the NVIDIA H100 Tensor Core GPU Transformer Engine and offers an easy-to-use and customizable Python interface.
NVIDIA Triton Inference Server software supports inference across cloud, data center, edge and embedded devices on GPUs, CPUs and other processors. Its integration can enable Ray developers to boost efficiency when deploying AI models from multiple deep learning and machine learning frameworks, including TensorRT, TensorFlow, PyTorch, ONNX, OpenVINO, Python, RAPIDS XGBoost and more.
With the NVIDIA NeMo framework, Ray users will be able to easily fine-tune and customize LLMs with business data, paving the way for LLMs that understand the unique offerings of individual businesses.
NeMo is an end-to-end, cloud-native framework to build, customize and deploy generative AI models anywhere. It features training and inferencing frameworks, guardrailing toolkits, data curation tools and pretrained models, offering enterprises an easy, cost-effective and fast way to adopt generative AI.
Options for Open-Source or Fully Supported Production AI Ray open source and the Anyscale Platform enable developers to effortlessly move from open source to deploying production AI at scale in the cloud.
The Anyscale Platform provides fully managed, enterprise-ready unified computing that makes it easy to build, deploy and manage scalable AI and Python applications using Ray, helping customers bring AI products to market faster at significantly lower cost.
Whether developers use Ray open source or the supported Anyscale Platform, Anyscale's core functionality helps them easily orchestrate LLM workloads. The NVIDIA AI integration can help developers build, train, tune and scale AI with even greater efficiency.
Ray and the Anyscale Platform run on accelerated computing from leading clouds, with the option to run on hybrid or multi-cloud computing. This helps developers easily scale up as they need more computing to power a successful LLM deployment.
The collaboration will also enable developers to begin building models on their workstations through NVIDIA AI Workbench and scale them easily across hybrid or multi-cloud accelerated computing once it's time to move to production.
NVIDIA AI integrations with Anyscale are in development and expected to be available by the end of the year.
Developers can sign up to get the latest news on this integration as well as a free 90-day evaluation of NVIDIA AI Enterprise.
To learn more, attend the Ray Summit in San Francisco this week or watch the demo video below.
See this notice regarding NVIDIA's software roadmap.
Most recent headlines
11/12/2025
Dalet, a leading provider of cloud-native, end-to-end media workflow solutions, ...
20/11/2025
LONDON The winners of the Rise Awards 2025, which recognize women and companies whose achievements have stood out in the media technology industry, have been an...
20/11/2025
SANTA MONICA & NEW YORK Lionsgate's Worldwide Television Distribution Group and Debmar-Mercury have launched MovieSphere Gold in more than 30 million homes....
20/11/2025
ALAMEDA, Calif. Clear-Com has introduced its four-channel HelixNet beltpack, a next-generation advancement of its widely used two-channel model, and anticipates...
20/11/2025
NEW PROVIDENCE, N.J. Streaming technology and services provider Kiswe has launched Kiswe Core, a next-generation cloud-based platform designed to simplify the s...
20/11/2025
MOUNTAIN VIEW, Calif. Dolby Laboratories has signed on as an official signature partner of the Bay Area Host Committee (BAHC) for the 2026 Super Bowl and the we...
20/11/2025
STUTTGART, Germany 3 Screen Solutions (3SS) has announced that the Canadian multiservice operator Telus has launched its Telus TV+ entertainment platform on Sam...
20/11/2025
Major League Baseball has announced deals with ESPN, NBCUniversal and Netflix covering 2026-2028 that make ESPN the exclusive rightsholder of MLB.TV, see regula...
20/11/2025
Berklee Presents Holiday Extravaganza: Duke Ellingtons The Nutcracker Suite and ...
19/11/2025
Sphere Sound Kicks Up Its Heels at Radio City Music HallUltra-immersive system from the singular Las Vegas venue is also ready for sportsBy Dan Daley, Audio Edi...
19/11/2025
The Future of Stadium and Arena Video Control Rooms: Tech Leaders Talk 4K, IP, a...
19/11/2025
One Year at Intuit Dome: Walking Through the Stunning Arena's Advanced Video...
19/11/2025
SVG Sit-Down: IMG's Francois Westcombe Goes Inside IMG's Annual Digital ...
19/11/2025
Inside Brighton & Hove Albion's fan-first media machine By George Bevir
Tuesday, November 18, 2025 - 10:10
Print This Story
Brighton & Hove Albion'...
19/11/2025
Rock Chalk Rebuild, Part 1: Kansas Athletics Brings New Production Power to Rebu...
19/11/2025
Rock Chalk Rebuild, Part 2: A Pro's Guide to Kansas Athletics' New Video...
19/11/2025
MLB Media Rights Shakeup: NBC Inks Three-Year Deal for Sunday Night Baseball, Pe...
19/11/2025
Rohan Parashuram Kanawade attends the premiere of Sabar Bonda (Cactus Pears) at the 2025 Sundance Film Festival at Egyptian Theatre on January 26, 2025, in Pa...
19/11/2025
Awards season is heating up, and the International Documentary Association just ...
19/11/2025
Aerojet Rocketdyne President Ken Bedingfield and Arkansas Gov. Sarah Huckabee Sanders were joined by state and local officials to celebrate the groundbreaking o...
19/11/2025
L3Harris and PentenAmio formalise their teaming agreement at MilCIS 2025, streng...
19/11/2025
Calrec gives Phoenix Broadcast Solutions its full support in strategic multi-year enterprise partnership New Singapore OB company invests in Calrec technology f...
19/11/2025
October brings a larger audience in front of TV screens. The cooler autumn weather and the continuation of new programming schedules brought increases in both t...
19/11/2025
IRVING, Texas Nexstar Media Group and Tegna filed applications on Nov. 18 with the Federal Communications Commission (FCC) seeking its consent to transfer broad...
19/11/2025
Vitec has acquired Datapath, a developer of real-time video processing for large-scale video walls, AVoIP content distribution and KVM control in mission-critic...
19/11/2025
NEVADA CITY, Calif. Telestream has introduced ARGUS Version 2.3, which introduces Live Look, a feature that lets operators inspect live and on-demand streams vi...
19/11/2025
STAMFORD, Conn. and NEW YORK Charter Communications, Inc. has announced a major deal with Amazon Web Services (AWS) that establishes AWS as one of Charters stra...
19/11/2025
TOKYO Atomos has announced a major firmware update that brings integrated camera control to the Ninja TX GO and Ninja TX its new CFexpress-based monitor-recor...
19/11/2025
Berklee's Signature Series Reimagines the Musical and Visual Artistry of Col...
19/11/2025
PlayBox Neo departed the NAB Show New York 2025 on a high note. The company's all-new PlayBox Neo Suite made a splash on the show floor, drawing strong inte...
19/11/2025
AAVBR (American Audio Visual Baton Rouge), a leading full-service event production company serving Louisiana and the Gulf Coast region, has enhanced its product...
19/11/2025
In its mission to deliver live event services combining advanced engineering with a creative approach to captivating live sports and entertainment experiences, ...
19/11/2025
NEW YORK Short-form vertical video has exploded across platforms like TikTok, Instagram and YouTube, but a new survey commissioned by Media.net, a provider of c...
19/11/2025
NEW YORK YES Network, the regional sports network home of the New York Yankees and Brooklyn Nets, and CAMB.AI have struck a partnership they said will leverage ...
19/11/2025
WASHINGTON The Federal Communications Commission has issued detailed guidance on filings that were disrupted by the government shutdown. While many of those fil...
19/11/2025
NEW YORK The popularity of NFL games is largely responsible for continued seasonal momentum in TV viewership into October (measured Sept. 29-Oct. 26), according...
19/11/2025
DirecTV has announced that its MyFree DirecTV streaming service has added seven new channels. Those include four sports services, NBA FAST, Red Bull TV, DAZN Ri...
19/11/2025
SURREY, U.K. Mark Roberts Motion Control (MRMC) has launched Flair Bridge, a compact device that transforms how operators control the company's motion-contr...
19/11/2025
WASHINGTON The Federal Communications Commission, which is continuing to rapidly ramp up its operations after the end of the federal government shutdown, has se...
19/11/2025
IRVING, Texas Nexstar Media Group and Tegna filed applications on Nov. 18 with the Federal Communications Commission (FCC) seeking its consent to transfer broad...
19/11/2025
Martin Freeman, Jessie Buckley, Wu-Tang Clan's Raekwon, Eddie the Eagle and ...
19/11/2025
Rohde & Schwarz collaborates with Broadcom to enable testing and validation of n...
19/11/2025
Back to All News
Innato, starring Imanol Arias and Elena Anaya, arrives on Netf...
19/11/2025
Back to All News
Guillermo del Toro Sits Down With Troll 2 Director Roar Uthaug...
19/11/2025
Some projects feel monumental from the moment the first phone call comes in. When BCE, the host broadcaster appointed to coordinate the national coverage, asked...
19/11/2025
November 19 2025, 08:00 (PST) BAY AREA HOST COMMITTEE AND DOLBY PARTNER TO DELIVER IMMERSIVE FAN EXPERIENCES
During the Week of Football's Biggest Game...
19/11/2025
Michael D. Higgins: Ireland's Ninth President features RT Archives footage and interviews across seven decades
Watch Wednesday 19 November at 9:35pm on RT...
19/11/2025
Scripps Research scientists receive $1.1 million to advance AI modeling for HIV vaccine development New AI system helps scientists rapidly pinpoint the most pro...
18/11/2025
Independent media across Central America are operating under intensifying financial pressure, yet there is a clear appetite for models that can sustain both ind...