Groups | Search | Server Info | Keyboard shortcuts | Login | Register [http] [https] [nntp] [nntps]
Groups > linux.kernel > #1542979
| From | Aaron Tomlin <atomlin@redhat.com> |
|---|---|
| Newsgroups | linux.kernel |
| Subject | Re: [PATCH] kernel/watchdog: Prevent false hardlockup on overloaded system |
| Date | 2016-12-15 19:50 +0100 |
| Message-ID | <sOJgm-PQ-27@gated-at.bofh.it> (permalink) |
| References | <sLqMV-4qH-11@gated-at.bofh.it> |
| Organization | linux.* mail to news gateway |
On Tue 2016-12-06 11:17 -0500, Don Zickus wrote: > On an overloaded system, it is possible that a change in the watchdog threshold > can be delayed long enough to trigger a false positive. > > This can easily be achieved by having a cpu spinning indefinitely on a task, > while another cpu updates watchdog threshold. > > What happens is while trying to park the watchdog threads, the hrtimers on the > other cpus trigger and reprogram themselves with the new slower watchdog > threshold. Meanwhile, the nmi watchdog is still programmed with the old faster > threshold. > > Because the one cpu is blocked, it prevents the thread parking on the other > cpus from completing, which is needed to shutdown the nmi watchdog and > reprogram it correctly. As a result, a false positive from the nmi watchdog is > reported. > > Fix this by setting a park_in_progress flag to block all lockups > until the parking is complete. > > Fix provided by Ulrich Obergfell. > > Cc: Ulrich Obergfell <uobergfe@redhat.com> > Signed-off-by: Don Zickus <dzickus@redhat.com> > --- > include/linux/nmi.h | 1 + > kernel/watchdog.c | 9 +++++++++ > kernel/watchdog_hld.c | 3 +++ > 3 files changed, 13 insertions(+) Looks fine to me. Reviewed-by: Aaron Tomlin <atomlin@redhat.com> -- Aaron Tomlin
Back to linux.kernel | Previous | Next | Find similar | Unroll thread
Re: [PATCH] kernel/watchdog: Prevent false hardlockup on overloaded system Aaron Tomlin <atomlin@redhat.com> - 2016-12-15 19:50 +0100
csiph-web