Groups | Search | Server Info | Keyboard shortcuts | Login | Register [http] [https] [nntp] [nntps]


Groups > linux.kernel > #1539746 > unrolled thread

Should xhci_irq() call usb_hc_died()?

Started byBjorn Helgaas <helgaas@kernel.org>
First post2016-12-10 01:30 +0100
Last post2016-12-12 11:50 +0100
Articles 3 — 3 participants

Back to article view | Back to linux.kernel


Contents

  Should xhci_irq() call usb_hc_died()? Bjorn Helgaas <helgaas@kernel.org> - 2016-12-10 01:30 +0100
    Re: Should xhci_irq() call usb_hc_died()? Felipe Balbi <felipe.balbi@linux.intel.com> - 2016-12-12 09:50 +0100
      Re: Should xhci_irq() call usb_hc_died()? Mathias Nyman <mathias.nyman@linux.intel.com> - 2016-12-12 11:50 +0100

#1539746 — Should xhci_irq() call usb_hc_died()?

FromBjorn Helgaas <helgaas@kernel.org>
Date2016-12-10 01:30 +0100
SubjectShould xhci_irq() call usb_hc_died()?
Message-ID<sMDI5-1Aj-11@gated-at.bofh.it>
Hi Mathias,

ehci_irq(), ohci_irq(), fotg210_irq(), and oxu210_hcd_irq() contain code
equivalent to this:

  status = ehci_readl(...);
  if (status == ~(u32) 0) {
    ...
    usb_hc_died(hcd);
    ...
    return IRQ_HANDLED;
  }

xhci_irq() has a similar check, but does not call usb_hc_died():

  status = readl(...);
  if (status = 0xffffffff) {
    ...
    return IRQ_HANDLED;
  }

Should xhci_irq() also call usb_hc_died()?  Maybe there's some reason
for it to be different than the others, but it wasn't obvious to this
casual observer :)

Bjorn

[toc] | [next] | [standalone]


#1540181

FromFelipe Balbi <felipe.balbi@linux.intel.com>
Date2016-12-12 09:50 +0100
Message-ID<sNut4-25U-7@gated-at.bofh.it>
In reply to#1539746

[Multipart message — attachments visible in raw view] — view raw

Hi,

Bjorn Helgaas <helgaas@kernel.org> writes:
> Hi Mathias,
>
> ehci_irq(), ohci_irq(), fotg210_irq(), and oxu210_hcd_irq() contain code
> equivalent to this:
>
>   status = ehci_readl(...);
>   if (status == ~(u32) 0) {
>     ...
>     usb_hc_died(hcd);
>     ...
>     return IRQ_HANDLED;
>   }
>
> xhci_irq() has a similar check, but does not call usb_hc_died():
>
>   status = readl(...);
>   if (status = 0xffffffff) {
>     ...
>     return IRQ_HANDLED;
>   }
>
> Should xhci_irq() also call usb_hc_died()?  Maybe there's some reason
> for it to be different than the others, but it wasn't obvious to this
> casual observer :)

you might just have fixed several bugs in dealing with a dead HC :-)

Can you provide a patch? (well, unless Mathias has a strong reason not
to call usb_hc_died(), of course).

-- 
balbi

[toc] | [prev] | [next] | [standalone]


#1540254

FromMathias Nyman <mathias.nyman@linux.intel.com>
Date2016-12-12 11:50 +0100
Message-ID<sNwlb-3eh-1@gated-at.bofh.it>
In reply to#1540181
On 12.12.2016 10:43, Felipe Balbi wrote:
>
> Hi,
>
> Bjorn Helgaas <helgaas@kernel.org> writes:
>> Hi Mathias,
>>
>> ehci_irq(), ohci_irq(), fotg210_irq(), and oxu210_hcd_irq() contain code
>> equivalent to this:
>>
>>    status = ehci_readl(...);
>>    if (status == ~(u32) 0) {
>>      ...
>>      usb_hc_died(hcd);
>>      ...
>>      return IRQ_HANDLED;
>>    }
>>
>> xhci_irq() has a similar check, but does not call usb_hc_died():
>>
>>    status = readl(...);
>>    if (status = 0xffffffff) {
>>      ...
>>      return IRQ_HANDLED;
>>    }
>>
>> Should xhci_irq() also call usb_hc_died()?  Maybe there's some reason
>> for it to be different than the others, but it wasn't obvious to this
>> casual observer :)

It probably should, I'm not aware of any reason why not, and a quick look at the
logs didn't reveal anything.

Currently we are calling usb_hcd_died() in a couple of timeout cases if we read
0xffffffff from the pci registers, So eventually usb_hc_died() will be called.

I'll take a look at this in more detail

>
> you might just have fixed several bugs in dealing with a dead HC :-)
>
> Can you provide a patch? (well, unless Mathias has a strong reason not
> to call usb_hc_died(), of course).

I don't think this is the worst case, there are a couple of other reasons such as
normal pci remove case we halt the host and reset the hardware after first HCD (USB2)
is removed, with all the secondary HCD (USB3) sand all its devices still connected,

Or then the abnormal case where HC disappears, we may time out while giving back a
killed URB, and  may end up never returning it. USB core waits with the roothub device
lock held for the URB, and we try to tear down xhci, which also requires the roothub
device lock at some point -> deadlock.

I'm am looking at these, but I need to make sure i fix it properly and not cause even
more issues.

-Mathias  
   

  

[toc] | [prev] | [standalone]


Back to top | Article view | linux.kernel


csiph-web