Merge pull request #51892 from xenolinux/haproxy-reload-interval

jeana-redhat · web-flow · commit dd820518be1c · 2022-11-11T08:56:24.000-05:00
OSDOCS-3937: Configure HAProxy reload interval
diff --git a/modules/configuring-haproxy-interval.adoc b/modules/configuring-haproxy-interval.adoc
@@ -0,0 +1,27 @@
+// Module included in the following assemblies:
+// * scalability_and_performance/routing-optimization.adoc
+// * post_installation_configuration/network-configuration.adoc
+
+:_content-type: PROCEDURE
+[id="configuring-haproxy-interval_{context}"]
+= Configuring HAProxy reload interval
+
+When you update a route or an endpoint associated with a route, {product-title} router updates the configuration for HAProxy. Then, HAProxy reloads the updated configuration for those changes to take effect. When HAProxy reloads, it generates a new process that handles new connections using the updated configuration.
+
+HAProxy keeps the old process running to handle existing connections until those connections are all closed. When old processes have long-lived connections, these processes can accumulate and consume resources.
+
+The default minimum HAProxy reload interval is five seconds. You can configure an Ingress Controller using its `spec.tuningOptions.reloadInterval` field to set a longer minimum reload interval.
+
+[WARNING]
+====
+Setting a large value for the minimum HAProxy reload interval can cause latency in observing updates to routes and their endpoints. To lessen the risk, avoid setting a value larger than the tolerable latency for updates.
+====
+
+.Procedure
+
+* Change the minimum HAProxy reload interval of the default Ingress Controller to 15 seconds by running the following command:
++
+[source, terminal]
+----
+$ oc -n openshift-ingress-operator patch ingresscontrollers/default --type=merge --patch='{"spec":{"tuningOptions":{"reloadInterval":"15s"}}}'
+----
diff --git a/modules/ingress-liveness-readiness-startup-probes.adoc b/modules/ingress-liveness-readiness-startup-probes.adoc
@@ -0,0 +1,48 @@
+// Module included in the following assemblies:
+// * scalability_and_performance/routing-optimization.adoc
+// * post_installation_configuration/network-configuration.adoc
+
+:_content-type: REFERENCE
+[id="ingress-liveness-readiness-startup-probes_{context}"]
+= Configuring Ingress Controller liveness, readiness, and startup probes
+
+Cluster administrators can configure the timeout values for the kubelet's liveness, readiness, and startup probes for router deployments that are managed by the {product-title} Ingress Controller (router). The liveness and readiness probes of the router use the default timeout value
+of 1 second, which is too short for the kubelet's probes to succeed in some scenarios. Probe timeouts can cause unwanted router restarts that interrupt application connections. The ability to set larger timeout values can reduce the risk of unnecessary and unwanted restarts.
+
+You can update the `timeoutSeconds` value on the `livenessProbe`, `readinessProbe`, and `startupProbe` parameters of the router container.
+
+[cols="3a,8a",options="header"]
+|===
+ |Parameter |Description
+
+ |`livenessProbe`
+ |The `livenessProbe` reports to the kubelet whether a pod is dead and needs to be restarted.
+
+ |`readinessProbe`
+ |The `readinessProbe` reports whether a pod is healthy or unhealthy. When the readiness probe reports an unhealthy pod, then the kubelet marks the pod as not ready to accept traffic. Subsequently, the endpoints for that pod are marked as not ready, and this status propogates to the kube-proxy. On cloud platforms with a configured load balancer, the kube-proxy communicates to the cloud load-balancer not to send traffic to the node with that pod.
+
+ |`startupProbe`
+ |The `startupProbe` gives the router pod up to 2 minutes to initialize before the kubelet begins sending the router liveness and readiness probes.  This initialization time can prevent routers with many routes or endpoints from prematurely restarting.
+|===
+
+
+[IMPORTANT]
+====
+The timeout configuration option is an advanced tuning technique that can be used to work around issues. However, these issues should eventually be diagnosed and possibly a support case or https://issues.redhat.com/secure/CreateIssueDetails!init.jspa?pid=12332330&summary=Summary&issuetype=1&priority=10200&versions=12385624[Jira issue] opened for any issues that causes probes to time out.
+====
+
+The following example demonstrates how you can directly patch the default router deployment to set a 5-second timeout for the liveness and readiness probes:
+
+
+[source, terminal]
+----
+$ oc -n openshift-ingress patch deploy/router-default --type=strategic --patch='{"spec":{"template":{"spec":{"containers":[{"name":"router","livenessProbe":{"timeoutSeconds":5},"readinessProbe":{"timeoutSeconds":5}}]}}}}'
+----
+
+.Verification
+[source, terminal]
+----
+$ oc -n openshift-ingress describe deploy/router-default | grep -e Liveness: -e Readiness:
+    Liveness:   http-get http://:1936/healthz delay=0s timeout=5s period=10s #success=1 #failure=3
+    Readiness:  http-get http://:1936/healthz/ready delay=0s timeout=5s period=10s #success=1 #failure=3
+----
diff --git a/modules/router-performance-optimizations.adoc b/modules/router-performance-optimizations.adoc
@@ -2,53 +2,10 @@
 // * scalability_and_performance/routing-optimization.adoc
 // * post_installation_configuration/network-configuration.adoc
 
-:_content-type: Procedure
+:_content-type: CONCEPT
 [id="router-performance-optimizations_{context}"]
 = Ingress Controller (router) performance optimizations
 
 {product-title} no longer supports modifying Ingress Controller deployments by setting environment variables such as `ROUTER_THREADS`, `ROUTER_DEFAULT_TUNNEL_TIMEOUT`, `ROUTER_DEFAULT_CLIENT_TIMEOUT`, `ROUTER_DEFAULT_SERVER_TIMEOUT`, and `RELOAD_INTERVAL`.
 
 You can modify the Ingress Controller deployment, but if the Ingress Operator is enabled, the configuration is overwritten.
-
-== Configuring Ingress Controller liveness, readiness, and startup probes
-
-Cluster administrators can configure the timeout values for the kubelet's liveness, readiness, and startup probes for router deployments that are managed by the {product-title} Ingress Controller (router). The liveness and readiness probes of the router use the default timeout value
-of 1 second, which is too short for the kubelet's probes to succeed in some scenarios. Probe timeouts can cause unwanted router restarts that interrupt application connections. The ability to set larger timeout values can reduce the risk of unnecessary and unwanted restarts.
-
-You can update the `timeoutSeconds` value on the `livenessProbe`, `readinessProbe`, and `startupProbe` parameters of the router container.
-
-[cols="3a,8a",options="header"]
-|===
- |Parameter |Description
-
- |`livenessProbe`
- |The `livenessProbe` reports to the kubelet whether a pod is dead and needs to be restarted.
-
- |`readinessProbe`
- |The `readinessProbe` reports whether a pod is healthy or unhealthy. When the readiness probe reports an unhealthy pod, then the kubelet marks the pod as not ready to accept traffic. Subsequently, the endpoints for that pod are marked as not ready, and this status propogates to the kube-proxy. On cloud platforms with a configured load balancer, the kube-proxy communicates to the cloud load-balancer not to send traffic to the node with that pod.
-
- |`startupProbe`
- |The `startupProbe` gives the router pod up to 2 minutes to initialize before the kubelet begins sending the router liveness and readiness probes.  This initialization time can prevent routers with many routes or endpoints from prematurely restarting.
-|===
-
-
-[IMPORTANT]
-====
-The timeout configuration option is an advanced tuning technique that can be used to work around issues. However, these issues should eventually be diagnosed and possibly a support case or https://issues.redhat.com/secure/CreateIssueDetails!init.jspa?pid=12332330&summary=Summary&issuetype=1&priority=10200&versions=12385624[Jira issue] opened for any issues that causes probes to time out.
-====
-
-The following example demonstrates how you can directly patch the default router deployment to set a 5-second timeout for the liveness and readiness probes:
-
-
-[source, terminal]
-----
-$ oc -n openshift-ingress patch deploy/router-default --type=strategic --patch='{"spec":{"template":{"spec":{"containers":[{"name":"router","livenessProbe":{"timeoutSeconds":5},"readinessProbe":{"timeoutSeconds":5}}]}}}}'
-----
-
-.Verification
-[source, terminal]
-----
-$ oc -n openshift-ingress describe deploy/router-default | grep -e Liveness: -e Readiness:
-    Liveness:   http-get http://:1936/healthz delay=0s timeout=5s period=10s #success=1 #failure=3
-    Readiness:  http-get http://:1936/healthz/ready delay=0s timeout=5s period=10s #success=1 #failure=3
-----
diff --git a/post_installation_configuration/network-configuration.adoc b/post_installation_configuration/network-configuration.adoc
@@ -111,6 +111,8 @@ The {product-title} HAProxy router scales to optimize performance.
 include::modules/baseline-router-performance.adoc[leveloffset=+2]
 
 include::modules/router-performance-optimizations.adoc[leveloffset=+2]
+include::modules/ingress-liveness-readiness-startup-probes.adoc[leveloffset=+3]
+include::modules/configuring-haproxy-interval.adoc[leveloffset=+3]
 
 [id="post-installation-osp-fips"]
 == Post-installation {rh-openstack} network configuration
@@ -121,4 +123,4 @@ include::modules/installation-osp-configuring-api-floating-ip.adoc[leveloffset=+
 include::modules/installation-osp-kuryr-port-pools.adoc[leveloffset=+2]
 include::modules/installation-osp-kuryr-settings-active.adoc[leveloffset=+2]
 include::modules/nw-osp-enabling-ovs-offload.adoc[leveloffset=+2]
-include::modules/nw-osp-hardware-offload-attaching-network.adoc[leveloffset=+2]
+include::modules/nw-osp-hardware-offload-attaching-network.adoc[leveloffset=+2]
diff --git a/scalability_and_performance/routing-optimization.adoc b/scalability_and_performance/routing-optimization.adoc
@@ -13,3 +13,5 @@ include::modules/baseline-router-performance.adoc[leveloffset=+1]
 For more information on Ingress sharding, see xref:../networking/ingress-operator.adoc#nw-ingress-sharding-route-labels_configuring-ingress[Configuring Ingress Controller sharding by using route labels] and xref:../networking/ingress-operator.adoc#nw-ingress-sharding-namespace-labels_configuring-ingress[Configuring Ingress Controller sharding by using namespace labels].
 
 include::modules/router-performance-optimizations.adoc[leveloffset=+1]
+include::modules/ingress-liveness-readiness-startup-probes.adoc[leveloffset=+2]
+include::modules/configuring-haproxy-interval.adoc[leveloffset=+2]