BERT
Model ● Covered in 14 stories + Follow
BERT is a transformer-based NLP model covered across multiple recent pieces focused on training and deploying BERT using Hugging Face tooling and hardware accelerators. Recent coverage includes tutorials for pre-training BERT on Habana Gaudi, fine-tuning BERT on Habana Gaudi for specific tasks, deploying BERT inference on AWS Inferentia, and running BERT as an IPU-ready model with Graphcore’s Optimum lineup. BERT is also referenced in broader transformer and framework updates, including expanded Transformer architecture documentation and discussions of Hugging Face’s TensorFlow integration (e.g., loading pretrained BERT models for Keras-based workflows).
Updated 15 September 2026
Specifications
No specifications recorded yet.
Latest developments
2023
2022
- Pre-Train BERT with Hugging Face Transformers and Habana Gaudi
- Hugging Face's TensorFlow Philosophy
- Graphcore and Hugging Face Launch New Lineup of IPU-Ready Transformers
- Getting Started with Transformers on Habana Gaudi
- Accelerate BERT inference with Hugging Face Transformers and AWS Inferentia
- BERT 101 - State Of The Art NLP Model Explained
2021
- Getting Started with Hugging Face Transformers for IPUs with Optimum
- Fine-Tune XLSR-Wav2Vec2 for low-resource ASR with 🤗 Transformers
- Scaling up BERT-like model Inference on modern CPU - Part 2
- Large Language Models: A New Moore's Law?
- How to Win a Data Hackathon (Hacklytics 2021)
2019
Relationships
Products & technology
- Hugging Face integrated with this model · 2 sources
- Google develops this model · 1 source
- Google integrated with this model · 1 source
- Graphcore integrated with this model · 1 source
- Intel integrated with this model · 1 source
- DistilBERT derived from this model · 1 source
- Hugging Face deploys this model · 1 source