WeAreDevelopers LIVE
•
May 26, 2021
Serverless deployment of (large) NLP models
Marek Suppa
- Discuss this with your agent
- Open in Claude
- Open in ChatGPT
Can you run massive NLP models within AWS Lambda's strict 250MB limit? Discover how Slido used knowledge distillation and ONNX to achieve sub-100ms serverless inference.
Matching moments
More from WeAreDevelopers LIVE
Related videos
Related articles
BB
Benedikt Bischof
BB
Benedikt Bischof
LM
Luis Minvielle
BB
Benedikt Bischof
From learning to earning
Jobs that call for the skills explored in this talk.
about 1 month ago
Machine Learning Engineer
Twilio
United States
Experienced
Remote
Spacy
Pytorch
Chatbots
about 1 month ago
•
Verified
LLM Training Engineer
Sciforium
San Francisco, United States
Expert
$155k–220k
Python
about 2 months ago
Machine Learning Engineer
TWILIO
Spain
Remote
Spacy
Twilio
Pytorch
3 months ago
Staff, Machine Learning Engineer (L4)
Twilio
Indianapolis, IN, United States
Expert
Remote
Keras
Presto
Twilio
about 1 month ago
Machine Learning Engineer
TWILIO
Cantabria, Spain
Remote
Twilio
Chatbots
Tensorflow
about 2 months ago
Machine Learning Engineer
TWILIO
Madrid, Spain
Remote
Twilio
Pytorch
Chatbots