Groups | Search | Server Info | Keyboard shortcuts | Login | Register [http] [https] [nntp] [nntps]


Groups > linux.kernel > #1632429 > unrolled thread

[PATCH 2/2] libnvdimm: clear region badblock in nvdimm_clear_poison()

Started byToshi Kani <toshi.kani@hpe.com>
First post2017-04-28 01:00 +0200
Last post2017-04-29 02:20 +0200
Articles 5 — 3 participants

Back to article view | Back to linux.kernel

This discussion starts older than the indexed window; earlier articles aren't shown. The article labeled Started by below is the oldest one visible, not the original post.


Contents

  [PATCH 2/2] libnvdimm: clear region badblock in nvdimm_clear_poison() Toshi Kani <toshi.kani@hpe.com> - 2017-04-28 01:00 +0200
    Re: [PATCH 2/2] libnvdimm: clear region badblock in nvdimm_clear_poison() Dan Williams <dan.j.williams@intel.com> - 2017-04-29 02:20 +0200
      Re: [PATCH 2/2] libnvdimm: clear region badblock in nvdimm_clear_poison() Dan Williams <dan.j.williams@intel.com> - 2017-04-29 02:40 +0200
        Re: [PATCH 2/2] libnvdimm: clear region badblock in  nvdimm_clear_poison() "Kani, Toshimitsu" <toshi.kani@hpe.com> - 2017-05-01 17:50 +0200
    Re: [PATCH 2/2] libnvdimm: clear region badblock in nvdimm_clear_poison() Dan Williams <dan.j.williams@intel.com> - 2017-04-29 02:20 +0200

#1632429 — [PATCH 2/2] libnvdimm: clear region badblock in nvdimm_clear_poison()

FromToshi Kani <toshi.kani@hpe.com>
Date2017-04-28 01:00 +0200
Subject[PATCH 2/2] libnvdimm: clear region badblock in nvdimm_clear_poison()
Message-ID<tB0ye-7ca-7@gated-at.bofh.it>
Badblocks are tracked at both region and device levels.
pmem_clear_poison() and nsio_rw_bytes() call nvdimm_clear_poison()
and then badblocks_clear() to clear badblocks at the device level.
However, it does not update badblocks at the region level, which
makes them inconsistent.

Change nvdimm_clear_poison() to update backblocks at the region
level to keep them consistent.

Signed-off-by: Toshi Kani <toshi.kani@hpe.com>
Cc: Dan Williams <dan.j.williams@intel.com>
Cc: Dave Jiang <dave.jiang@intel.com>
Cc: Vishal Verma <vishal.l.verma@intel.com>
---
Based on 'libnvdimm-for-next'.
---
 drivers/nvdimm/bus.c |    9 +++++++++
 1 file changed, 9 insertions(+)

diff --git a/drivers/nvdimm/bus.c b/drivers/nvdimm/bus.c
index 43ddfd4..998332d 100644
--- a/drivers/nvdimm/bus.c
+++ b/drivers/nvdimm/bus.c
@@ -179,6 +179,7 @@ long nvdimm_clear_poison(struct device *dev, phys_addr_t phys,
 	struct nvdimm_bus_descriptor *nd_desc;
 	struct nd_cmd_clear_error clear_err;
 	struct nd_cmd_ars_cap ars_cap;
+	struct resource res;
 	u32 clear_err_unit, mask;
 	int cmd_rc, rc;
 
@@ -222,6 +223,14 @@ long nvdimm_clear_poison(struct device *dev, phys_addr_t phys,
 	if (clear_err.cleared > 0)
 		nvdimm_forget_poison(nvdimm_bus, phys, clear_err.cleared);
 
+	if (clear_err.cleared > 0 && clear_err.cleared / 512) {
+		nvdimm_bus_lock(&nvdimm_bus->dev);
+		res.start = phys;
+		res.end = phys + clear_err.cleared - 1;
+		__nvdimm_bus_badblocks_clear(nvdimm_bus, &res);
+		nvdimm_bus_unlock(&nvdimm_bus->dev);
+	}
+
 	return clear_err.cleared;
 }
 EXPORT_SYMBOL_GPL(nvdimm_clear_poison);

[toc] | [next] | [standalone]


#1633210

FromDan Williams <dan.j.williams@intel.com>
Date2017-04-29 02:20 +0200
Message-ID<tBohb-6MR-1@gated-at.bofh.it>
In reply to#1632429
On Fri, Apr 28, 2017 at 5:10 PM, Dan Williams <dan.j.williams@intel.com> wrote:
> On Thu, Apr 27, 2017 at 3:57 PM, Toshi Kani <toshi.kani@hpe.com> wrote:
>> Badblocks are tracked at both region and device levels.
>> pmem_clear_poison() and nsio_rw_bytes() call nvdimm_clear_poison()
>> and then badblocks_clear() to clear badblocks at the device level.
>> However, it does not update badblocks at the region level, which
>> makes them inconsistent.
>>
>> Change nvdimm_clear_poison() to update backblocks at the region
>> level to keep them consistent.
>>
>> Signed-off-by: Toshi Kani <toshi.kani@hpe.com>
>> Cc: Dan Williams <dan.j.williams@intel.com>
>> Cc: Dave Jiang <dave.jiang@intel.com>
>> Cc: Vishal Verma <vishal.l.verma@intel.com>
>
> This looks good, and it seems we have a bug in the other location that
> does this in __nd_ioctl(). That other one is missing the
> "clear_err.cleared / 512" check. Can you respin this and define a
> common helper that both locations can call?

On second thought, I'll take this and spin my own cleanup / fix on top.

Thanks Toshi!

[toc] | [prev] | [next] | [standalone]


#1633214

FromDan Williams <dan.j.williams@intel.com>
Date2017-04-29 02:40 +0200
Message-ID<tBoAx-6WJ-3@gated-at.bofh.it>
In reply to#1633210
On Fri, Apr 28, 2017 at 5:12 PM, Dan Williams <dan.j.williams@intel.com> wrote:
> On Fri, Apr 28, 2017 at 5:10 PM, Dan Williams <dan.j.williams@intel.com> wrote:
>> On Thu, Apr 27, 2017 at 3:57 PM, Toshi Kani <toshi.kani@hpe.com> wrote:
>>> Badblocks are tracked at both region and device levels.
>>> pmem_clear_poison() and nsio_rw_bytes() call nvdimm_clear_poison()
>>> and then badblocks_clear() to clear badblocks at the device level.
>>> However, it does not update badblocks at the region level, which
>>> makes them inconsistent.
>>>
>>> Change nvdimm_clear_poison() to update backblocks at the region
>>> level to keep them consistent.
>>>
>>> Signed-off-by: Toshi Kani <toshi.kani@hpe.com>
>>> Cc: Dan Williams <dan.j.williams@intel.com>
>>> Cc: Dave Jiang <dave.jiang@intel.com>
>>> Cc: Vishal Verma <vishal.l.verma@intel.com>
>>
>> This looks good, and it seems we have a bug in the other location that
>> does this in __nd_ioctl(). That other one is missing the
>> "clear_err.cleared / 512" check. Can you respin this and define a
>> common helper that both locations can call?
>
> On second thought, I'll take this and spin my own cleanup / fix on top.
>
> Thanks Toshi!

...and I need to drop it again because it causes this:

[  106.974889] BUG: sleeping function called from invalid context at
kernel/locking/mutex.c:74
7
[  106.977328] in_atomic(): 1, irqs_disabled(): 0, pid: 5584, name: dd
[  106.978845] 1 lock held by dd/5584:
[  106.979923]  #0:  (&bdev->bd_mutex){+.+.+.}, at:
[<ffffffff812f4937>] __blkdev_put+0x47/0x3
70
[  106.982221] CPU: 29 PID: 5584 Comm: dd Tainted: G           O
4.11.0-rc4+ #105
[  106.984329] Hardware name: QEMU Standard PC (i440FX + PIIX, 1996),
BIOS 1.9.3-1.fc25 04/01/
2014
[  106.986638] Call Trace:
[  106.987530]  dump_stack+0x86/0xc3
[  106.988555]  ___might_sleep+0x17d/0x250
[  106.989650]  __might_sleep+0x4a/0x80
[  106.990718]  __mutex_lock+0x58/0x980
[  106.991788]  ? nvdimm_bus_lock+0x21/0x30 [libnvdimm]
[  106.993059]  ? _raw_spin_unlock+0x27/0x40
[  106.994181]  ? debug_lockdep_rcu_enabled+0x1d/0x20
[  106.995430]  mutex_lock_nested+0x1b/0x20
[  106.996552]  nvdimm_bus_lock+0x21/0x30 [libnvdimm]
[  106.997804]  nvdimm_clear_poison+0x11a/0x150 [libnvdimm]
[  106.999138]  nsio_rw_bytes+0x18f/0x280 [libnvdimm]
[  107.000390]  btt_write_pg+0x1d4/0x3c0 [nd_btt]

[toc] | [prev] | [next] | [standalone]


#1633705 — Re: [PATCH 2/2] libnvdimm: clear region badblock in nvdimm_clear_poison()

From"Kani, Toshimitsu" <toshi.kani@hpe.com>
Date2017-05-01 17:50 +0200
SubjectRe: [PATCH 2/2] libnvdimm: clear region badblock in nvdimm_clear_poison()
Message-ID<tClKi-2vD-17@gated-at.bofh.it>
In reply to#1633214
On Fri, 2017-04-28 at 17:35 -0700, Dan Williams wrote:
> On Fri, Apr 28, 2017 at 5:12 PM, Dan Williams <dan.j.williams@intel.c
> om> wrote:
> > On Fri, Apr 28, 2017 at 5:10 PM, Dan Williams <dan.j.williams@intel
> > .com> wrote:
> > > On Thu, Apr 27, 2017 at 3:57 PM, Toshi Kani <toshi.kani@hpe.com>
> > > wrote:
> > > > Badblocks are tracked at both region and device levels.
> > > > pmem_clear_poison() and nsio_rw_bytes() call
> > > > nvdimm_clear_poison() and then badblocks_clear() to clear
> > > > badblocks at the device level. However, it does not update
> > > > badblocks at the region level, which makes them inconsistent.
> > > > 
> > > > Change nvdimm_clear_poison() to update backblocks at the region
> > > > level to keep them consistent.
> > > > 
> > > > Signed-off-by: Toshi Kani <toshi.kani@hpe.com>
> > > > Cc: Dan Williams <dan.j.williams@intel.com>
> > > > Cc: Dave Jiang <dave.jiang@intel.com>
> > > > Cc: Vishal Verma <vishal.l.verma@intel.com>
> > > 
> > > This looks good, and it seems we have a bug in the other location
> > > that does this in __nd_ioctl(). That other one is missing the
> > > "clear_err.cleared / 512" check. Can you respin this and define a
> > > common helper that both locations can call?
> > 
> > On second thought, I'll take this and spin my own cleanup / fix on
> > top.
> > 
> > Thanks Toshi!
> 
> ...and I need to drop it again because it causes this:
> 
> [  106.974889] BUG: sleeping function called from invalid context at
> kernel/locking/mutex.c:747
> [  106.977328] in_atomic(): 1, irqs_disabled(): 0, pid: 5584, name:
> dd

Sorry, I should have tested with CONFIG_DEBUG_ATOMIC_SLEEP set.  Thanks
for straighten it out in your patch! 

-Toshi 

[toc] | [prev] | [next] | [standalone]


#1633213

FromDan Williams <dan.j.williams@intel.com>
Date2017-04-29 02:20 +0200
Message-ID<tBohb-6MR-3@gated-at.bofh.it>
In reply to#1632429
On Thu, Apr 27, 2017 at 3:57 PM, Toshi Kani <toshi.kani@hpe.com> wrote:
> Badblocks are tracked at both region and device levels.
> pmem_clear_poison() and nsio_rw_bytes() call nvdimm_clear_poison()
> and then badblocks_clear() to clear badblocks at the device level.
> However, it does not update badblocks at the region level, which
> makes them inconsistent.
>
> Change nvdimm_clear_poison() to update backblocks at the region
> level to keep them consistent.
>
> Signed-off-by: Toshi Kani <toshi.kani@hpe.com>
> Cc: Dan Williams <dan.j.williams@intel.com>
> Cc: Dave Jiang <dave.jiang@intel.com>
> Cc: Vishal Verma <vishal.l.verma@intel.com>

This looks good, and it seems we have a bug in the other location that
does this in __nd_ioctl(). That other one is missing the
"clear_err.cleared / 512" check. Can you respin this and define a
common helper that both locations can call?

[toc] | [prev] | [standalone]


Back to top | Article view | linux.kernel


csiph-web