⚡|serverless

2030 threads · Page 9 of 41

Why is pod availability so low? 4 messages
New version of worker-vllm ? 2 messages
Whisper X 8 messages
GPU Memory Management for Concurrent Request 4 messages
Load Balancer
Serverless spins up needless workers during cold boot 9 messages
4/20 healthy workers - pruna ai image model from hub 3 messages
Queue
Increase serverless quotas 3 messages
GitHub builds for serverless stuck in "Pending" 2 messages
how can I upgrade vllm version in serverless worker image? 4 messages
Is anyone there? Runpod's service has failed. 2 messages
Queue
Is runpod down? 6 messages
401 Error 12 messages
Can't send request to serverless: 10 messages
GOT 401 Error 8 messages
Failed to trigger build 2 messages
Sending Image File to Serverless LLAMA3.2 Vision 2 messages
Serveless prices and technical questions 30 messages
Solved
value not in list on serverless 6 messages
Solved Model Store
Finish task with error: CUDA error: no kernel image is available for execution on the device 3 messages
Runpod underwater? 8 messages
Failing requests 39 messages
A deposit error caused me to lose money, any Runpod staff providing support? 3 messages
So whats the deal with all the issues and the charging for failed docker fetches? 11 messages
Queue Load Balancer
Having trouble using ModelPatchLoader in Comfyui Serverless 5 messages
I have some questions about Serverless scaling and account limits 5 messages
GPU Detection Failure Across 20–50% of Workers — Months of Unresolved Issues 18 messages
Question about Serverless max workers, account quota limits, and using multiple accounts 2 messages
How to use Python package for public endpoints? 3 messages
Are there video cards for the workers that are functioning correctly now? 2 messages
Finish task with error: CUDA error: no kernel image is available for execution on the device 2 messages
Throttling on multiple endpoints and failed workers 12 messages
Ongoing Throttling Issues with Multiple Serverless Endpoints 4 messages
Serverless throttled 42 messages
Solved
Huggingface cached models seems to not working 39 messages
vLLM jobs not processing: "deferring container creation" 25 messages
Jobs seem to be stuck in a queue(loadbalancer ep) - workers are running but not processing requests 7 messages
Queue Load Balancer
Serverless Worker Crashed but Request Still Running 30 messages
/dev/nvidia-caps never mounts 5 messages
Nvidia-smi parsing error 7 messages
Workflow TO Api wizard not working properly 3 messages
RuntimeError: CUDA driver initialization failed, you might not have a CUDA gpu. 5 messages
The Delay Time is extremely long 2 messages
Serverless loadbalancer scaling config (Request Count) 6 messages
Serverless FAILING to add Workers 53 messages
Serverless crashing 11 messages
Workers and rate increasing 17 messages
Queue
what am i doing wrong, serverless workers optimization 20 messages
"ComfyUI to API"-wizard 45 messages
Stuck initializing vLLM 26 messages
Get a cost with the result? 2 messages