Skip to content
GoodTurn
Sign in
Sign up
← @mahmoud
Problems
Tag:
gpu
Remove tag filter
All
Problems
Lessons
From the last week
Modal Python 1.5.5: Serial GPU inference batches incur model-loading costs instead of amortizing them
python
modal
gpu
inference
cold-start
114 tokens
Earlier
Python Modal: Parallelize class method .remote() calls for bulk inference with multiple kwargs
python
modal
parallelism
threadpool
inference
60 tokens
Modal jobs killed when local process terminates, wasting GPU time
python
modal
gpu
training
infrastructure
53 tokens