Difference between revisions of "Benchmarks"

From
Jump to: navigation, search
Line 20: Line 20:
  
 
<youtube>YygGzfkhtJc</youtube>
 
<youtube>YygGzfkhtJc</youtube>
<youtube>hQRBLW6giRc</youtube>
 
 
<youtube>WlXhpXv9kDU</youtube>
 
<youtube>WlXhpXv9kDU</youtube>
 
<youtube>wpQiEHYkBys</youtube>
 
<youtube>wpQiEHYkBys</youtube>
 +
<youtube>lgK0BlXdOCw</youtube>
  
  
Line 34: Line 34:
  
 
<youtube>uz_eYqutEG4</youtube>
 
<youtube>uz_eYqutEG4</youtube>
 +
 +
 +
== MLPerf ==
 +
* [http://mlperf.org/ MLPerf] benchmarks for measuring training and inference performance of ML hardware, software, and services.
 +
 +
<youtube>MKII0KXDqn4</youtube>
 +
<youtube>hQRBLW6giRc</youtube>
 +
<youtube>0Fuxjq1eiZ4</youtube>
 +
<youtube>sH03-InVba4</youtube>

Revision as of 13:01, 24 December 2019

YouTube search... ...Google search


GLUE

The General Language Understanding Evaluation (GLUE) benchmark is a collection of resources for training, evaluating, and analyzing natural language understanding systems. GLUE consists of: A benchmark of nine sentence- or sentence-pair language understanding tasks built on established existing datasets and selected to cover a diverse range of dataset sizes, text genres, and degrees of difficulty, A diagnostic dataset designed to evaluate and analyze model performance with respect to a wide range of linguistic phenomena found in natural language, and A public leaderboard for tracking performance on the benchmark and a dashboard for visualizing the performance of models on the diagnostic set.


MLPerf

  • MLPerf benchmarks for measuring training and inference performance of ML hardware, software, and services.