Scaling-up BERT Inference on CPU (Part 1)
Hugging Face Blog
· huggingface.co · April 20, 2021
Read on Hugging Face Blog →
Opens the original article in a new tab.
More from the blog
Microsoft Research Blog
· microsoft.com
Aug 31, 2026
GigaPath-Flash and GigaTIME-Flash: Toward population-scale discovery with efficient pathology foundation models
What if pathology foundation models could do more with less? GigaPath-Flash and GigaTIME-Flash cut computational demands while maintaining strong performance, opening the door to larger studies and broader exploration. The post GigaPath-Flash and GigaTIME-Flash: Toward population-scale discovery with efficient pathology foundation models appeared first on Microsoft Research.
Google DeepMind Blog
· deepmind.google
Aug 21, 2026
From Atari to EVE Online: Building on 15 Years of AI Research in Games
Google DeepMind partners with game studios to prototype breakthrough AI gameplay.
Hugging Face Blog
· huggingface.co
Aug 21, 2026
How Hugging Face Inference Endpoints, Jobs, and Buckets Power Search on Papers with Code
Hugging Face Blog
· huggingface.co
Aug 20, 2026
Up to 3.2x Faster Inference with LFM2.5-DSpark
A Blog post by Liquid AI on Hugging Face