Note: This bug is displayed in read-only format because the product is no longer active in Red Hat Bugzilla.

Bug 1811342

Summary: Host unreachable during loopback test
Product: OpenShift Container Platform Reporter: Clayton Coleman <ccoleman>
Component: NetworkingAssignee: Jacob Tanenbaum <jtanenba>
Networking sub component: openshift-sdn QA Contact: zhaozhanqi <zzhao>
Status: CLOSED DUPLICATE Docs Contact:
Severity: high    
Priority: unspecified CC: aconstan, anbhat, bbennett, jtanenba
Version: 4.4   
Target Milestone: ---   
Target Release: 4.4.0   
Hardware: Unspecified   
OS: Unspecified   
Whiteboard:
Fixed In Version: Doc Type: If docs needed, set a value
Doc Text:
Story Points: ---
Clone Of: Environment:
Last Closed: 2020-03-11 13:56:15 UTC Type: Bug
Regression: --- Mount Type: ---
Documentation: --- CRM:
Verified Versions: Category: ---
oVirt Team: --- RHEL 7.3 requirements from Atomic Host:
Cloudforms Team: --- Target Upstream Version:
Embargoed:

Description Clayton Coleman 2020-03-07 21:10:59 UTC
Host unreachable in a 4.4 release job, this should not be happening since this is supposed to be a loopback test on the host itself.

The test itself does not appear to be flaky, but rather there appears to be some lower level disruption on the node.  This is a release blocker (evidence of lower level regression) unless we can triage the issue to something fairly scoped.

[Area:Networking] services basic functionality [Top Level] [Area:Networking] services basic functionality should allow connections to another pod on the same node via a service IP [Suite:openshift/conformance/parallel] expand_less	45s
fail [github.com/openshift/origin/test/extended/networking/services.go:16]: Expected success, but got an error:
    <exec.CodeExitError>: {
        Err: {
            s: "error running /usr/bin/kubectl --server=https://api.ci-op-ik7ngs0x-91171.origin-ci-int-gce.dev.openshift.com:6443 --kubeconfig=/tmp/admin.kubeconfig exec --namespace=e2e-net-services1-3416 execpod-sourceip-ci-op-rkrm8-w-b-ljv8h.c.openshift-gce-devq4dn2 -- /bin/sh -x -c wget -T 30 -qO- 172.30.14.21:8080:\nCommand stdout:\n\nstderr:\n+ wget -T 30 -qO- 172.30.14.21:8080\nwget: can't connect to remote host (172.30.14.21): Host is unreachable\ncommand terminated with exit code 1\n\nerror:\nexit status 1",
        },
        Code: 1,
    }
    error running /usr/bin/kubectl --server=https://api.ci-op-ik7ngs0x-91171.origin-ci-int-gce.dev.openshift.com:6443 --kubeconfig=/tmp/admin.kubeconfig exec --namespace=e2e-net-services1-3416 execpod-sourceip-ci-op-rkrm8-w-b-ljv8h.c.openshift-gce-devq4dn2 -- /bin/sh -x -c wget -T 30 -qO- 172.30.14.21:8080:
    Command stdout:
    
    stderr:
    + wget -T 30 -qO- 172.30.14.21:8080
    wget: can't connect to remote host (172.30.14.21): Host is unreachable
    command terminated with exit code 1
    
    error:
    exit status 1

Comment 2 Clayton Coleman 2020-03-07 21:14:36 UTC
Another similar failure but this time across nodes

https://prow.svc.ci.openshift.org/view/gcs/origin-ci-test/logs/release-openshift-origin-installer-e2e-gcp-4.4/1924

[Area:Networking] services basic functionality [Top Level] [Area:Networking] services basic functionality should allow connections to another pod on a different node via a service IP [Suite:openshift/conformance/parallel] expand_less	40s
fail [github.com/openshift/origin/test/extended/networking/services.go:20]: Expected success, but got an error:
    <exec.CodeExitError>: {
        Err: {
            s: "error running /usr/bin/kubectl --server=https://api.ci-op-5q7s8fk7-91171.origin-ci-int-gce.dev.openshift.com:6443 --kubeconfig=/tmp/admin.kubeconfig exec --namespace=e2e-net-services1-2393 execpod-sourceip-ci-op-ftcsq-w-c-n9r4j.c.openshift-gce-devt488h -- /bin/sh -x -c wget -T 30 -qO- 172.30.231.33:8080:\nCommand stdout:\n\nstderr:\n+ wget -T 30 -qO- 172.30.231.33:8080\nwget: can't connect to remote host (172.30.231.33): Host is unreachable\ncommand terminated with exit code 1\n\nerror:\nexit status 1",
        },
        Code: 1,
    }
    error running /usr/bin/kubectl --server=https://api.ci-op-5q7s8fk7-91171.origin-ci-int-gce.dev.openshift.com:6443 --kubeconfig=/tmp/admin.kubeconfig exec --namespace=e2e-net-services1-2393 execpod-sourceip-ci-op-ftcsq-w-c-n9r4j.c.openshift-gce-devt488h -- /bin/sh -x -c wget -T 30 -qO- 172.30.231.33:8080:
    Command stdout:
    
    stderr:
    + wget -T 30 -qO- 172.30.231.33:8080
    wget: can't connect to remote host (172.30.231.33): Host is unreachable
    command terminated with exit code 1
    
    error:
    exit status 1

Comment 6 Ben Bennett 2020-03-11 13:56:15 UTC
We are seeing the same segfaults from iptables-save and iptables-restore as with https://bugzilla.redhat.com/show_bug.cgi?id=1812261

*** This bug has been marked as a duplicate of bug 1812261 ***