News
Newest
Ask
Show
Jobs
Open on GitHub
Ask HN: Slow OpenAI Inference on AWS Bedrock
Inference on AWS Bedrock for OpenAI models has been very slow and glitchy the last few days with today being the worst. Very long latency of over 5 minutes to first token. What's going on?
2 points | by
timedude
18 hours ago
2 comments
mikert89
3 hours ago
likely they sold all the gpu capacity to enterprise clients
timedude
18 hours ago
us-east-1 seems to be the worst region right now.
2 comments