Home / Companies / Lambda / Blog / Post Details
Content Deep Dive

NVIDIA Hopper: H100 and FP8 Support

Blog post from Lambda

Post Details
Company
Date Published
Author
Jeremy Hummel
Word Count
1,245
Company Posts That Month
2
Language
English
Hacker News Points
-
Post removed?
No
Summary

NVIDIA's H100 Tensor Core GPU introduces native support for FP8 data types, which offer a significant increase in delivered application performance by 2x and reduce memory requirements by 2x compared to 16-bit floating-point. The proposed FP8 specification includes two formats: E4M3 and E5M2, with varying ranges and precision. FP8 is particularly useful for reducing memory requirements, allowing the training of larger models or decreasing training time, which can result in significant cost savings on cloud-based GPU usage. Additionally, FP8 inference can offer up to 4.5x speedup compared to previous results, while also retaining a higher accuracy compared to INT8 quantization methods like post-training quantization and quantization-aware training.

Trends Found in this Post
Trend Post Mentions Total Month Mentions Posts Companies MoM
Serverless 1 601 116 65 -58%
Use This Data

Use this post, company, and trend context to find content marketing opportunities, perform competitive analysis, or address product feature gaps via the Plushcap MCP server or the Plushcap API.