频繁的工人超时

我已经设置了gunicorn与3个工人，30个工人连接和使用eventlet工人类。它被设置在Nginx后面。每请求几次，我就会在日志里看到这个。

[ERROR] gunicorn.error: WORKER TIMEOUT (pid:23475)
None
[INFO] gunicorn.error: Booting worker with pid: 23514

为什么会这样?我怎样才能知道哪里出了问题呢?

当前回答

检查你的工人没有被健康检查杀死。长请求可能会阻塞健康检查请求，worker会被平台杀死，因为平台认为worker没有响应。

例如，如果您有一个25秒长的请求，并且活动检查被配置为每10秒命中同一服务中的不同端点，1秒超时，并重试3次，这就给出了10+1*3 ~ 13秒，您可以看到它会触发一些时间，但并不总是如此。

如果是这种情况，解决方案是重新配置您的活动检查(或您的平台使用的任何健康检查机制)，以便它可以等待您的典型请求完成。或者允许更多的线程——这样可以确保健康检查不会阻塞足够长的时间来触发worker kill。

你可以看到，增加更多的工人可能有助于(或隐藏)这个问题。

2022-10-10 16:01:28

其他回答

这招对我很管用:

gunicorn app:app -b :8080 --timeout 120 --workers=3 --threads=3 --worker-connections=1000

如果你有eventlet，添加:

--worker-class=eventlet

如果你有gevent添加:

--worker-class=gevent

2020-06-08 01:01:23

使用——log-level debug运行Gunicorn。

它应该会给你一个应用程序堆栈跟踪。

2012-08-18 16:21:42

WORKER TIMEOUT表示应用程序不能在规定的时间内响应请求。你可以使用gunicorn超时设置来设置。一些应用程序需要比另一个应用程序更多的时间来响应。

另一个可能影响这一点的因素是员工类型的选择

The default synchronous workers assume that your application is resource-bound in terms of CPU and network bandwidth. Generally this means that your application shouldn’t do anything that takes an undefined amount of time. An example of something that takes an undefined amount of time is a request to the internet. At some point the external network will fail in such a way that clients will pile up on your servers. So, in this sense, any web application which makes outgoing requests to APIs will benefit from an asynchronous worker.

当我遇到与您相同的问题时(我试图使用Docker Swarm部署我的应用程序)，我尝试增加超时并使用另一种类型的工人类。但都失败了。

然后我突然意识到我的资源限制太低在我的撰写文件中的服务。在我的例子中，这就是减慢应用程序的原因

deploy:
  replicas: 5
  resources:
    limits:
      cpus: "0.1"
      memory: 50M
  restart_policy:
    condition: on-failure

所以我建议你先检查一下是什么减慢了你的应用程序

2018-08-09 09:31:47

除了已经建议的gunicorn超时设置，因为你在前面使用nginx，你可以检查这两个参数是否有效，proxy_connect_timeout和proxy_read_timeout默认为60秒。可以在nginx配置文件中这样设置它们，

proxy_connect_timeout 120s;
proxy_read_timeout 120s;

2023-01-14 17:32:58

超时是这个问题的一个关键参数。

然而，它不适合我。

当我设置workers=1时，我发现没有gunicorn超时错误。

当我看我的代码，我发现一些套接字连接(套接字。在服务器init中发送& socket.recv)。

套接字。Recv将阻塞我的代码，这就是为什么它总是超时时，工人>1

希望能给那些对我有意见的人一些建议

2019-11-28 08:53:04

频繁的工人超时

推荐文章

最新文章

标签