NVIDIA ACE official logo, the Avatar Cloud Engine suite of generative AI microservices for digital humans, used by ServiceNow, Dell, and Perfect World Games

Bring digital humans to life with generative AI microservices for speech, intelligence, and animation

Not yet rated
Released 2022
Visit Website

Gallery

4 items

NVIDIA ACE video thumbnail
VIDEO
NVIDIA ACE screenshot 2
2
NVIDIA ACE screenshot 3
3
NVIDIA ACE screenshot 4
4

About NVIDIA ACE

NVIDIA ACE (Avatar Cloud Engine) is a suite of generative AI microservices designed to bring lifelike, interactive digital humans and avatars to games, customer service applications, and enterprise software. NVIDIA first announced the technology as the Omniverse Avatar Cloud Engine on August 9, 2022, positioning it as a set of cloud-native AI models that would let businesses of any size access the computing power needed to build assistants and avatars capable of understanding multiple languages, responding to speech, and making intelligent recommendations. The platform reached general availability for cloud deployment in mid-2024, alongside early access for local RTX AI PCs, with adoption from companies spanning customer service, gaming, and healthcare, including Dell Technologies, ServiceNow, Inventec, and Perfect World Games.

At its technical core, ACE combines several specialized NVIDIA AI microservices that developers can mix and match: NVIDIA Riva handles automatic speech recognition, text-to-speech synthesis, and translation across multiple languages; NVIDIA Nemotron large language models provide contextual language understanding and response generation, including Nemotron-3 4.5B, a small language model purpose-built to run on-device with accuracy comparable to cloud-based LLMs; and NVIDIA Audio2Face generates realistic facial animation directly from an audio track, with a companion Audio2Emotion model that infers a character's emotional state from speech to drive more expressive performances. Because these are packaged as modular NVIDIA NIM (NVIDIA Inference Microservice) containers, developers can run them flexibly across the cloud or locally on RTX-powered PCs depending on available hardware.

ACE's clearest showcase to date is Covert Protocol, a technology demo built by Inworld AI in partnership with NVIDIA and unveiled at GDC and GTC 2024, where players take on the role of a detective having real, unscripted conversations with AI-driven digital human characters built on Unreal Engine 5 and MetaHuman. Inworld's engine layers cognition, perception, and behavior systems on top of ACE's speech and animation microservices, letting each conversation and playthrough unfold differently based on what the player actually says, a meaningfully different approach to non-player characters than pre-scripted dialogue trees. Beyond gaming, ACE technologies are being adopted for telehealth digital humans, enterprise customer service avatars, and virtual factory applications involving humanoid robots.

Pricing follows NVIDIA's broader NIM and AI Enterprise licensing model rather than a simple per-seat or per-token structure: hosted NIM microservices are free to call for prototyping and evaluation through build.nvidia.com and the NVIDIA Developer Program, subject to model-specific rate limits, while production deployment requires an NVIDIA AI Enterprise license at $4,500 per GPU per year (or approximately $1 per GPU-hour in the cloud), with a free 90-day production-grade evaluation license available to test before committing. This makes NVIDIA ACE best suited for game studios, enterprise software teams, and healthcare or customer service platforms with real engineering resources and NVIDIA GPU infrastructure who want to build genuinely interactive, unscripted digital humans, rather than teams looking for a simple, low-cost off-the-shelf avatar widget.

Key Features

  • NVIDIA Riva for automatic speech recognition, text-to-speech, and translation
  • NVIDIA Nemotron LLMs for language understanding and contextual responses
  • NVIDIA Audio2Face for real-time facial animation generated from audio
  • Audio2Emotion models for inferring emotional state from speech
  • Flexible deployment across cloud, on-premises, or local RTX AI PCs
  • NVIDIA AI Inference Manager SDK for simplified PC deployment
  • Packaged as NVIDIA NIM microservices for containerized, portable deployment
  • Unreal Engine plugin support for direct integration into game development pipelines

Pros

  • Modular microservices architecture lets developers mix only the components (speech, language, animation) they actually need
  • Free prototyping tier on build.nvidia.com makes it genuinely accessible to test before any purchase commitment
  • Flexible deployment across cloud, on-premises DGX, or local RTX AI PCs avoids locking teams into one infrastructure path
  • Backed by NVIDIA's dominant GPU ecosystem, giving strong performance guarantees on certified hardware
  • Real, publicly showcased use cases like Covert Protocol demonstrate genuinely unscripted, real-time AI conversation, not just a tech demo claim

Cons

  • No public list price for production use; AI Enterprise licensing requires contacting NVIDIA sales or a partner for a quote
  • Production deployment realistically requires NVIDIA GPU hardware and the $4,500/GPU/year AI Enterprise license, a real cost barrier for smaller teams
  • Best-fit use cases (gaming NPCs, enterprise digital humans) require meaningful engineering investment to integrate multiple microservices together
  • Free tier is genuinely limited to prototyping and evaluation on build.nvidia.com, not production traffic
  • Workstation vs. server licensing distinctions (e.g., on DGX Spark) can create ambiguity for teams unsure which license tier actually applies to their deployment

Pricing

NVIDIA ACE microservices follow NVIDIA's NIM (NVIDIA Inference Microservice) pricing model rather than a standalone product price. Hosted NIM endpoints are free to call for prototyping and evaluation through build.nvidia.com under the NVIDIA Developer Program, subject to model- and account-specific rate limits that NVIDIA does not publish as one universal quota. NIM containers can also be downloaded for free from NGC for development, testing, and experimentation on up to 16 GPUs. Production deployment requires an NVIDIA AI Enterprise license, priced at $4,500 per GPU per year for a self-managed subscription (or roughly $1 per GPU-hour when billed through cloud marketplaces on AWS, Azure, or GCP), with a $22,500 perpetual license option also referenced in third-party analysis. A free 90-day production-grade evaluation license is available before committing to a paid subscription. Public list pricing for the full AI Enterprise suite is not published; exact costs depend on GPU type, deployment scale (cloud vs. on-premises DGX), and support tier (Business Standard vs. Business Critical 24x7 coverage), and require contacting NVIDIA sales or an authorized partner for a formal quote.

Claim Verified Creator Badge

Are you the founder of NVIDIA ACE? Display this listing's verified badge on your website to show your customers that your product has been vetted and listed on AI Central Resources.

FEATURED ONAI Central Resources
HTML Embed Code
<a href="https://www.aicentralresources.com/tool/nvidia-ace" target="_blank" rel="noopener">
  <img src="https://www.aicentralresources.com/badges/featured-badge-dark.svg" alt="Featured on AICentralResources" width="200" height="54" style="border: none;" />
</a>

* Place this HTML snippet in your website's footer, landing page, or press section. This creates a search-friendly backlink directly to your verification page.

Connect with NVIDIA ACE

Frequently Asked Questions

NVIDIA ACE (Avatar Cloud Engine) is a suite of generative AI microservices for building and deploying interactive digital humans and avatars, combining speech recognition, language understanding, facial animation, and emotion inference for use cases spanning gaming NPCs, customer service agents, and enterprise digital assistants.

NVIDIA ACE microservices are free to access for prototyping and evaluation through build.nvidia.com and the NVIDIA Developer Program, with model- and account-specific rate limits. Production deployment requires an NVIDIA AI Enterprise license, priced at $4,500 per GPU per year (or roughly $1 per GPU-hour in the cloud), with a free 90-day production-grade evaluation license also available before committing.

NVIDIA ACE's core components include NVIDIA Riva for automatic speech recognition, text-to-speech, and translation; NVIDIA Nemotron LLMs for language understanding and contextual responses; and NVIDIA Audio2Face for generating realistic facial animation directly from an audio track.

NVIDIA ACE was first announced as the Omniverse Avatar Cloud Engine on August 9, 2022, positioned as a suite of cloud-native AI models for building lifelike virtual assistants and digital humans. It reached general availability for the cloud in mid-2024, with early access for RTX AI PCs shortly after.

Yes. NVIDIA ACE technologies power the Covert Protocol tech demo built by Inworld AI, showcased at GDC and GTC 2024, where players act as a detective interacting with AI-driven digital human characters in real time. Companies including Dell Technologies, ServiceNow, Perfect World Games, and Hippocratic AI have also adopted ACE for customer service, gaming, and healthcare applications.

Similar AI Tools to NVIDIA ACE

View all alternatives of NVIDIA ACE
Custom
Charisma.ai

Charisma.ai

Immersive conversational AI for training simulations and brand experiences

Not yet rated
Freemium
Convai

Convai

Build lifelike 3D AI characters with real-time vision, voice, and memory for games and virtual worlds

Not yet rated
Freemium
DeepMotion

DeepMotion

Generate 3D animations from video in seconds with AI motion capture.

Not yet rated
Freemium
ElevenLabs

ElevenLabs

Lifelike AI voices, agents, and audio for creators and developers

Not yet rated
Freemium
HeyGen

HeyGen

Transform ideas into stunning AI videos with realistic avatars and multi-language support.

Not yet rated
Freemium
Inworld AI

Inworld AI

Realtime TTS, STT, and LLM routing infrastructure for consumer-scale voice AI

Not yet rated