Bug 2019886

Summary: Kuryr unable to finish ports recovery upon controller restart
Product: OpenShift Container Platform Reporter: Maysa Macedo <mdemaced>
Component: NetworkingAssignee: Maysa Macedo <mdemaced>
Networking sub component: kuryr QA Contact: Itzik Brown <itbrown>
Status: CLOSED ERRATA Docs Contact:
Severity: high    
Priority: high CC: itbrown, mdulko
Version: 4.9   
Target Milestone: ---   
Target Release: 4.10.0   
Hardware: Unspecified   
OS: Unspecified   
Whiteboard:
Fixed In Version: Doc Type: No Doc Update
Doc Text:
Story Points: ---
Clone Of: Environment:
Last Closed: 2022-03-10 16:24:56 UTC Type: Bug
Regression: --- Mount Type: ---
Documentation: --- CRM:
Verified Versions: Category: ---
oVirt Team: --- RHEL 7.3 requirements from Atomic Host:
Cloudforms Team: --- Target Upstream Version:
Embargoed:
Bug Depends On:    
Bug Blocks: 2022684    

Description Maysa Macedo 2021-11-03 14:36:49 UTC
Description of problem:

When there is a big number of Pods on the cluster and the controller restarts, the pre-created ports need to be recovered and the pools populated again, this operation consumes a considerable amount of time, which makes the kuryr-controller to restart as is not able to handle other requests like Namespace deletion and pools clean up the pools due to pools not fully populated.

Version-Release number of selected component (if applicable):
OCP 4.9

How reproducible:


Steps to Reproduce:
1.
2.
3.

Actual results:


Expected results:


Additional info:

Comment 3 Itzik Brown 2021-11-29 13:00:34 UTC
Created 5 deploymentd in 5 different namespaces each with 100 pods
Restarted Kuryr controller 
It took 10s for the controller to be ready

Version
OCP 4.10.0-0.nightly-2021-11-27-004934
OSP RHOS-16.1-RHEL-8-20210903.n.0

Comment 5 ShiftStack Bugwatcher 2022-03-05 07:07:22 UTC
Removing the Triaged keyword because:
* the QE automation assessment (flag qe_test_coverage) is missing

Comment 7 errata-xmlrpc 2022-03-10 16:24:56 UTC
Since the problem described in this bug report should be
resolved in a recent advisory, it has been closed with a
resolution of ERRATA.

For information on the advisory (Moderate: OpenShift Container Platform 4.10.3 security update), and where to find the updated
files, follow the link below.

If the solution does not work for you, open a new bug report.

https://access.redhat.com/errata/RHSA-2022:0056