{"id":129649,"date":"2024-03-27T15:40:47","date_gmt":"2024-03-27T15:40:47","guid":{"rendered":"https:\/\/blogs.nvidia.com\/blog\/tensorrt-llm-inference-mlperf\/"},"modified":"2024-03-27T15:40:47","modified_gmt":"2024-03-27T15:40:47","slug":"nvidia-hopper-leaps-ahead-in-generative-ai-at-mlperf","status":"publish","type":"post","link":"https:\/\/webdomino.net\/index.php\/nvidia\/nvidia-hopper-leaps-ahead-in-generative-ai-at-mlperf\/","title":{"rendered":"NVIDIA Hopper Leaps Ahead in Generative AI at MLPerf"},"content":{"rendered":"It\u2019s official: NVIDIA delivered the world\u2019s fastest platform in industry-standard tests for inference on generative AI. In the latest MLPerf benchmarks, NVIDIA TensorRT-LLM \u2014 software that speeds and simplifies the complex job of inference on large language models \u2014 boosted the performance of NVIDIA Hopper architecture GPUs on the GPT-J LLM nearly 3x over their\t\t<a class=\"read-more\" href=\"https:\/\/blogs.nvidia.com\/blog\/tensorrt-llm-inference-mlperf\/\">\n\t\t\tRead Article\t\t\t<span data-icon=\"y\"><\/span>\n\t\t<\/a>\n\t","protected":false},"excerpt":{"rendered":"<p>It\u2019s official: NVIDIA delivered the world\u2019s fastest platform in industry-standard tests for inference on generative AI. In the latest MLPerf benchmarks, NVIDIA TensorRT-LLM \u2014 software that speeds and simplifies the complex job of inference on large lan&#8230;<\/p>\n","protected":false},"author":266,"featured_media":0,"comment_status":"closed","ping_status":"closed","sticky":false,"template":"","format":"standard","meta":{"_mi_skip_tracking":false,"_monsterinsights_sitenote_active":false,"_monsterinsights_sitenote_note":"","_monsterinsights_sitenote_category":0,"footnotes":""},"categories":[343],"tags":[],"yoast_head":"<!-- This site is optimized with the Yoast SEO plugin v21.6 - https:\/\/yoast.com\/wordpress\/plugins\/seo\/ -->\n<title>NVIDIA Hopper Leaps Ahead in Generative AI at MLPerf - WebDomino.NET<\/title>\n<meta name=\"robots\" content=\"index, follow, max-snippet:-1, max-image-preview:large, max-video-preview:-1\" \/>\n<link rel=\"canonical\" href=\"https:\/\/webdomino.net\/index.php\/nvidia\/nvidia-hopper-leaps-ahead-in-generative-ai-at-mlperf\/\" \/>\n<meta property=\"og:locale\" content=\"en_US\" \/>\n<meta property=\"og:type\" content=\"article\" \/>\n<meta property=\"og:title\" content=\"NVIDIA Hopper Leaps Ahead in Generative AI at MLPerf - WebDomino.NET\" \/>\n<meta property=\"og:description\" content=\"It\u2019s official: NVIDIA delivered the world\u2019s fastest platform in industry-standard tests for inference on generative AI. In the latest MLPerf benchmarks, NVIDIA TensorRT-LLM \u2014 software that speeds and simplifies the complex job of inference on large lan...\" \/>\n<meta property=\"og:url\" content=\"https:\/\/webdomino.net\/index.php\/nvidia\/nvidia-hopper-leaps-ahead-in-generative-ai-at-mlperf\/\" \/>\n<meta property=\"og:site_name\" content=\"WebDomino.NET\" \/>\n<meta property=\"article:published_time\" content=\"2024-03-27T15:40:47+00:00\" \/>\n<meta name=\"author\" content=\"Dave Salvator\" \/>\n<meta name=\"twitter:label1\" content=\"Written by\" \/>\n\t<meta name=\"twitter:data1\" content=\"Dave Salvator\" \/>\n<script type=\"application\/ld+json\" class=\"yoast-schema-graph\">{\"@context\":\"https:\/\/schema.org\",\"@graph\":[{\"@type\":\"WebPage\",\"@id\":\"https:\/\/webdomino.net\/index.php\/nvidia\/nvidia-hopper-leaps-ahead-in-generative-ai-at-mlperf\/\",\"url\":\"https:\/\/webdomino.net\/index.php\/nvidia\/nvidia-hopper-leaps-ahead-in-generative-ai-at-mlperf\/\",\"name\":\"NVIDIA Hopper Leaps Ahead in Generative AI at MLPerf - WebDomino.NET\",\"isPartOf\":{\"@id\":\"https:\/\/webdomino.net\/#website\"},\"datePublished\":\"2024-03-27T15:40:47+00:00\",\"dateModified\":\"2024-03-27T15:40:47+00:00\",\"author\":{\"@id\":\"https:\/\/webdomino.net\/#\/schema\/person\/310dae90f82d7d270510022746aa300c\"},\"breadcrumb\":{\"@id\":\"https:\/\/webdomino.net\/index.php\/nvidia\/nvidia-hopper-leaps-ahead-in-generative-ai-at-mlperf\/#breadcrumb\"},\"inLanguage\":\"en-US\",\"potentialAction\":[{\"@type\":\"ReadAction\",\"target\":[\"https:\/\/webdomino.net\/index.php\/nvidia\/nvidia-hopper-leaps-ahead-in-generative-ai-at-mlperf\/\"]}]},{\"@type\":\"BreadcrumbList\",\"@id\":\"https:\/\/webdomino.net\/index.php\/nvidia\/nvidia-hopper-leaps-ahead-in-generative-ai-at-mlperf\/#breadcrumb\",\"itemListElement\":[{\"@type\":\"ListItem\",\"position\":1,\"name\":\"Home\",\"item\":\"https:\/\/webdomino.net\/\"},{\"@type\":\"ListItem\",\"position\":2,\"name\":\"NVIDIA Hopper Leaps Ahead in Generative AI at MLPerf\"}]},{\"@type\":\"WebSite\",\"@id\":\"https:\/\/webdomino.net\/#website\",\"url\":\"https:\/\/webdomino.net\/\",\"name\":\"WebDomino.NET\",\"description\":\"Global NEWS WordPress site\",\"potentialAction\":[{\"@type\":\"SearchAction\",\"target\":{\"@type\":\"EntryPoint\",\"urlTemplate\":\"https:\/\/webdomino.net\/?s={search_term_string}\"},\"query-input\":\"required name=search_term_string\"}],\"inLanguage\":\"en-US\"},{\"@type\":\"Person\",\"@id\":\"https:\/\/webdomino.net\/#\/schema\/person\/310dae90f82d7d270510022746aa300c\",\"name\":\"Dave Salvator\",\"image\":{\"@type\":\"ImageObject\",\"inLanguage\":\"en-US\",\"@id\":\"https:\/\/webdomino.net\/#\/schema\/person\/image\/\",\"url\":\"https:\/\/secure.gravatar.com\/avatar\/60146dd4624401222bdbdfebf7cd1c66?s=96&d=mm&r=g\",\"contentUrl\":\"https:\/\/secure.gravatar.com\/avatar\/60146dd4624401222bdbdfebf7cd1c66?s=96&d=mm&r=g\",\"caption\":\"Dave Salvator\"},\"sameAs\":[\"https:\/\/nvidianews.nvidia.com\"],\"url\":\"https:\/\/webdomino.net\/index.php\/author\/dave-salvator\/\"}]}<\/script>\n<!-- \/ Yoast SEO plugin. -->","yoast_head_json":{"title":"NVIDIA Hopper Leaps Ahead in Generative AI at MLPerf - WebDomino.NET","robots":{"index":"index","follow":"follow","max-snippet":"max-snippet:-1","max-image-preview":"max-image-preview:large","max-video-preview":"max-video-preview:-1"},"canonical":"https:\/\/webdomino.net\/index.php\/nvidia\/nvidia-hopper-leaps-ahead-in-generative-ai-at-mlperf\/","og_locale":"en_US","og_type":"article","og_title":"NVIDIA Hopper Leaps Ahead in Generative AI at MLPerf - WebDomino.NET","og_description":"It\u2019s official: NVIDIA delivered the world\u2019s fastest platform in industry-standard tests for inference on generative AI. In the latest MLPerf benchmarks, NVIDIA TensorRT-LLM \u2014 software that speeds and simplifies the complex job of inference on large lan...","og_url":"https:\/\/webdomino.net\/index.php\/nvidia\/nvidia-hopper-leaps-ahead-in-generative-ai-at-mlperf\/","og_site_name":"WebDomino.NET","article_published_time":"2024-03-27T15:40:47+00:00","author":"Dave Salvator","twitter_misc":{"Written by":"Dave Salvator"},"schema":{"@context":"https:\/\/schema.org","@graph":[{"@type":"WebPage","@id":"https:\/\/webdomino.net\/index.php\/nvidia\/nvidia-hopper-leaps-ahead-in-generative-ai-at-mlperf\/","url":"https:\/\/webdomino.net\/index.php\/nvidia\/nvidia-hopper-leaps-ahead-in-generative-ai-at-mlperf\/","name":"NVIDIA Hopper Leaps Ahead in Generative AI at MLPerf - WebDomino.NET","isPartOf":{"@id":"https:\/\/webdomino.net\/#website"},"datePublished":"2024-03-27T15:40:47+00:00","dateModified":"2024-03-27T15:40:47+00:00","author":{"@id":"https:\/\/webdomino.net\/#\/schema\/person\/310dae90f82d7d270510022746aa300c"},"breadcrumb":{"@id":"https:\/\/webdomino.net\/index.php\/nvidia\/nvidia-hopper-leaps-ahead-in-generative-ai-at-mlperf\/#breadcrumb"},"inLanguage":"en-US","potentialAction":[{"@type":"ReadAction","target":["https:\/\/webdomino.net\/index.php\/nvidia\/nvidia-hopper-leaps-ahead-in-generative-ai-at-mlperf\/"]}]},{"@type":"BreadcrumbList","@id":"https:\/\/webdomino.net\/index.php\/nvidia\/nvidia-hopper-leaps-ahead-in-generative-ai-at-mlperf\/#breadcrumb","itemListElement":[{"@type":"ListItem","position":1,"name":"Home","item":"https:\/\/webdomino.net\/"},{"@type":"ListItem","position":2,"name":"NVIDIA Hopper Leaps Ahead in Generative AI at MLPerf"}]},{"@type":"WebSite","@id":"https:\/\/webdomino.net\/#website","url":"https:\/\/webdomino.net\/","name":"WebDomino.NET","description":"Global NEWS WordPress site","potentialAction":[{"@type":"SearchAction","target":{"@type":"EntryPoint","urlTemplate":"https:\/\/webdomino.net\/?s={search_term_string}"},"query-input":"required name=search_term_string"}],"inLanguage":"en-US"},{"@type":"Person","@id":"https:\/\/webdomino.net\/#\/schema\/person\/310dae90f82d7d270510022746aa300c","name":"Dave Salvator","image":{"@type":"ImageObject","inLanguage":"en-US","@id":"https:\/\/webdomino.net\/#\/schema\/person\/image\/","url":"https:\/\/secure.gravatar.com\/avatar\/60146dd4624401222bdbdfebf7cd1c66?s=96&d=mm&r=g","contentUrl":"https:\/\/secure.gravatar.com\/avatar\/60146dd4624401222bdbdfebf7cd1c66?s=96&d=mm&r=g","caption":"Dave Salvator"},"sameAs":["https:\/\/nvidianews.nvidia.com"],"url":"https:\/\/webdomino.net\/index.php\/author\/dave-salvator\/"}]}},"_links":{"self":[{"href":"https:\/\/webdomino.net\/index.php\/wp-json\/wp\/v2\/posts\/129649"}],"collection":[{"href":"https:\/\/webdomino.net\/index.php\/wp-json\/wp\/v2\/posts"}],"about":[{"href":"https:\/\/webdomino.net\/index.php\/wp-json\/wp\/v2\/types\/post"}],"author":[{"embeddable":true,"href":"https:\/\/webdomino.net\/index.php\/wp-json\/wp\/v2\/users\/266"}],"replies":[{"embeddable":true,"href":"https:\/\/webdomino.net\/index.php\/wp-json\/wp\/v2\/comments?post=129649"}],"version-history":[{"count":1,"href":"https:\/\/webdomino.net\/index.php\/wp-json\/wp\/v2\/posts\/129649\/revisions"}],"predecessor-version":[{"id":129650,"href":"https:\/\/webdomino.net\/index.php\/wp-json\/wp\/v2\/posts\/129649\/revisions\/129650"}],"wp:attachment":[{"href":"https:\/\/webdomino.net\/index.php\/wp-json\/wp\/v2\/media?parent=129649"}],"wp:term":[{"taxonomy":"category","embeddable":true,"href":"https:\/\/webdomino.net\/index.php\/wp-json\/wp\/v2\/categories?post=129649"},{"taxonomy":"post_tag","embeddable":true,"href":"https:\/\/webdomino.net\/index.php\/wp-json\/wp\/v2\/tags?post=129649"}],"curies":[{"name":"wp","href":"https:\/\/api.w.org\/{rel}","templated":true}]}}