BERT
Model ● Covered in 14 stories + Follow
BERT is a transformer-based NLP model covered across multiple recent pieces focused on training and deploying BERT using Hugging Face tooling and hardware accelerators. Recent coverage includes tutorials for pre-training BERT on Habana Gaudi, fine-tuning BERT on Habana Gaudi for specific tasks, deploying BERT inference on AWS Inferentia, and running BERT as an IPU-ready model with Graphcore’s Optimum lineup. BERT is also referenced in broader transformer and framework updates, including expanded Transformer architecture documentation and discussions of Hugging Face’s TensorFlow integration (e.g., loading pretrained BERT models for Keras-based workflows).
Updated 15 September 2026
Specifications
No specifications recorded yet.
Latest developments
Q4 2023
Q1 2023
Q3 2022
Q2 2022
- Graphcore and Hugging Face Launch New Lineup of IPU-Ready Transformers
- Getting Started with Transformers on Habana Gaudi
Q1 2022
- Accelerate BERT inference with Hugging Face Transformers and AWS Inferentia
- BERT 101 - State Of The Art NLP Model Explained
Q4 2021
- Getting Started with Hugging Face Transformers for IPUs with Optimum
- Fine-Tune XLSR-Wav2Vec2 for low-resource ASR with 🤗 Transformers
- Scaling up BERT-like model Inference on modern CPU - Part 2
- Large Language Models: A New Moore's Law?
Q1 2021
Q1 2019
Relationships
Products & technology
- Hugging Face integrated with this model · 2 sources
- Google develops this model · 1 source
- Google integrated with this model · 1 source
- Graphcore integrated with this model · 1 source
- Intel integrated with this model · 1 source
- DistilBERT derived from this model · 1 source
- Hugging Face deploys this model · 1 source