โ Back to all news
Hugging Face Blog
Accelerate BERT inference with Hugging Face Transformers and AWS Inferentia
Related stories
Hugging Face Blog
Feb 1, 2024
Hugging Face Text Generation Inference available for AWS Inferentia2
Hugging Face Blog
Aug 22, 2022
Pre-Train BERT with Hugging Face Transformers and Habana Gaudi
Hugging Face Blog
Aug 6
Baseten on Hugging Face Inference Providers ๐ฅ
Hugging Face Blog
Oct 24, 2023
Deploy Embedding Models with Hugging Face Inference Endpoints
Hugging Face Blog
Jun 16, 2025
Groq on Hugging Face Inference Providers ๐ฅ
Hugging Face Blog
Nov 24, 2025
OVHcloud on Hugging Face Inference Providers ๐ฅ
Hugging Face Blog
Apr 29
DeepInfra on Hugging Face Inference Providers ๐ฅ
Hugging Face Blog
Nov 4, 2021
Scaling up BERT-like model Inference on modern CPU - Part 2
Hugging Face Blog
Apr 16, 2025
Cohere on Hugging Face Inference Providers ๐ฅ
Hugging Face Blog
Apr 17, 2023
Accelerating Hugging Face Transformers with AWS Inferentia2
Hugging Face Blog
Apr 20, 2021
Scaling-up BERT Inference on CPU (Part 1)
Hugging Face Blog
Oct 14, 2022
Getting Started with Hugging Face Inference Endpoints