3031 Tisch Way, 110 Plaza West
San Jose, CA 95128

©2016 by Chambiz Inc.

Nvidia breaks records in training and inference for real-time conversational AI

August 20, 2019

Nvidia’s GPU-powered platform for developing and running conversational AI that understands and responds to natural language requests has achieved some key milestones and broken some records that have big implications for anyone building on their tech – which includes companies large and small, since much of the code they’ve used to achieve these advancements is open source, written in PyTorch and easy to run.

 

The biggest achievements Nvidia announced include its breaking the hour mark in training BERT, one of the world’s most advanced AI language models and a state-of-the-art model widely considered a good standard for natural language processing. Nvidia’s AI platform was able to train the model in under an hour, a record-breaking achievement at just 53 minutes, and the trained model could then successfully infer (ie, actually applying the learned capability achieved through training to achieve results) in under 2 milliseconds (10 milliseconds is considered a high-water mark in the industry), another record.

 

Nvidia’s breakthroughs aren’t just cause for bragging rights – these advances scale and provide real-world benefits for anyone working with their NLP conversational AI and GPU hardware. Nvidia achieved its record-setting times for training on one of its SuperPOD systems which is made up of 92 Nvidia DGX-2H systems runnings 1,472 V100 GPUs, and managed the inference on Nvidia T4 GPUs running Nvidia TensorRT – which beat the performance of event highly optimized CPUs by many orders of magnitude. But it’s making available the BERT training code, and TensorRT optimized BERT Sample via GitHub  for all to leverage.

 


Alongside these milestones, Nvidia’s Research wing also build and trained the largest ever language model based on ‘Transformers,’ which is the tech that underlies BERT, too. This custom model includes a massive 8.3 billion parameters, making it 24 times the size of BERT-Large, the largest current core BERT model. Nvidia has cheekily titled this model ‘Megatron,’ and also offered up the PyTorch code it used to train this model so that others can also train their own similar, massive Transformer-based language models.

 

 

 

Please reload

Our Recent Posts

We are at the beginning of a new era, where biology has shifted from an empirical science to an engineering discipline. After a millennia of using man...

Biology is Eating the World: A Manifesto

November 9, 2019

The Q3 2019 Global Venture Capital Report: Seed Stage Deals Increase While Broader Funding Environment Shows Signs Of Erosion

October 13, 2019

From The Shadows Of The Fortune 500, Atlanta Emerges As A Tech Hub

October 5, 2019

1/1
Please reload

Tags

Please reload