Groups | Search | Server Info | Keyboard shortcuts | Login | Register [http] [https] [nntp] [nntps]


Groups > linux.kernel > #1735849

[PATCH 7/7] fs-writeback: only allow one inflight and pending full flush

Path csiph.com!news.redatomik.org!weretis.net!feeder4.news.weretis.net!storethat.news.telefonica.de!telefonica.de!news.panservice.it!bofh.it!news.nic.it!robomod
From Jens Axboe <axboe@kernel.dk>
Newsgroups linux.kernel
Subject [PATCH 7/7] fs-writeback: only allow one inflight and pending full flush
Date Wed, 20 Sep 2017 17:40:02 +0200
Message-ID <urPgu-4TY-19@gated-at.bofh.it> (permalink)
References <urPgt-4TY-13@gated-at.bofh.it>
X-Original-To linux-kernel@vger.kernel.org, linux-fsdevel@vger.kernel.org, linux-mm@kvack.org
Dkim-Signature v=1; a=rsa-sha256; c=relaxed/relaxed; d=kernel-dk.20150623.gappssmtp.com; s=20150623; h=from:to:cc:subject:date:message-id:in-reply-to:references; bh=I8oerl0v/43bfeTxmLnVxPD7d59YErQnIu6Toz2a/8s=; b=kMg4+R66uHkJtwUrpxIffvcXETQLEu3k6M0fJtf/xGCgMKpBy7AJ6uyRaZXMkE7rpe 6psfXjLQ8/+jSml3lHRJK5096HeldUrYXm2GLF1Hg1XnsnlzPY7jDS6AsBQB79hUjsBE RwXNaHTAWpazUEpxwuUM6+2lX6XYk7ViypJJipWCkiIMlginVLvNrA0JOOTTA2ROhNgc GzJ0K8FRx16LUKA0bZA30ocIC2zYFQ/gzQRGSp7lELNlPtG4W3rmQJW6O7+3aR6TLAH2 klMKhC6BOroHSwL9IHV80MKpQuOpWDIW0yVl2TFXP6SKxXZ1qKdsk4G7TcyQGXJ6aq3Q eaaQ==
X-Google-Dkim-Signature v=1; a=rsa-sha256; c=relaxed/relaxed; d=1e100.net; s=20161025; h=x-gm-message-state:from:to:cc:subject:date:message-id:in-reply-to :references; bh=I8oerl0v/43bfeTxmLnVxPD7d59YErQnIu6Toz2a/8s=; b=jG6AeQdqAE+9n2w+YecPnOa/8GA41XjqjBm4L8+ISAVRJR36I8Ts8cDSkzzJaELiY0 wMUD3UcqxHbMeqS2NfITHllBbFX7Nj4jB8E4BSE1QlWHSr4lQNX0HCQIneyitkSa164a nbdIethZucgiYuYC7/7/D9MVI5XFHh7y7UKoZGx4UgNgGLhmVuOYYHjthlinxqjgUpLn a4X29T1Axjd/rN3ZlES5BjK4L9PQZqgxOuSABBvfYJDM0len7QGm/rEWhFfiVpDwcpX6 rydm/BVSolieyaSUyRndNpnIo3jP1lFhicPNYHhdQvNqt7OJzVN/8xSiFWwfXeeAAvqK VwjQ==
X-Gm-Message-State AHPjjUjvlYnHPY6T7l81gu3Azx3jF7Sv0zZNIBdKca1eIELOyKDoPF9G cNLWfJJpeFnSvSuKFy9ZuzXDj1UICfM=
X-Google-SMTP-Source AOwi7QAF4TiXlxtWDsRBRWr3VEqlIKoeDHXRTn0PqbxIo00LoWyPeXoebL2SJEWbOzeJ6Sz+1osJMg==
X-Received by 10.36.189.14 with SMTP id x14mr3369780ite.30.1505921609679; Wed, 20 Sep 2017 08:33:29 -0700 (PDT)
X-Mailer git-send-email 2.7.4
Sender robomod@news.nic.it
List-ID <linux-kernel.vger.kernel.org>
X-Mailing-List linux-kernel@vger.kernel.org
Approved robomod@news.nic.it
Lines 104
Organization linux.* mail to news gateway
X-Original-Cc hannes@cmpxchg.org, clm@fb.com, jack@suse.cz, Jens Axboe <axboe@kernel.dk>
X-Original-Date Wed, 20 Sep 2017 09:33:02 -0600
X-Original-Message-ID <1505921582-26709-8-git-send-email-axboe@kernel.dk>
X-Original-References <1505921582-26709-1-git-send-email-axboe@kernel.dk>
X-Original-Sender linux-kernel-owner@vger.kernel.org
Xref csiph.com linux.kernel:1735849

Show key headers only | View raw


When someone calls wakeup_flusher_threads() or
wakeup_flusher_threads_bdi(), they schedule writeback of all dirty
pages in the system (or on that bdi). If we are tight on memory, we
can get tons of these queued from kswapd/vmscan. This causes (at
least) two problems:

1) We consume a ton of memory just allocating writeback work items.
2) We spend so much time processing these work items, that we
   introduce a softlockup in writeback processing.

Fix this by adding a 'start_all' bit to the writeback structure, and
set that when someone attempts to flush all dirty page.  The bit is
cleared when we start writeback on that work item. If the bit is
already set when we attempt to queue !nr_pages writeback, then we
simply ignore it.

This provides us one full flush in flight, with one pending as well,
and makes for more efficient handling of this type of writeback.

Acked-by: Johannes Weiner <hannes@cmpxchg.org>
Tested-by: Chris Mason <clm@fb.com>
Reviewed-by: Jan Kara <jack@suse.cz>
Signed-off-by: Jens Axboe <axboe@kernel.dk>
---
 fs/fs-writeback.c                | 24 ++++++++++++++++++++++++
 include/linux/backing-dev-defs.h |  1 +
 2 files changed, 25 insertions(+)

diff --git a/fs/fs-writeback.c b/fs/fs-writeback.c
index 3916ea2484ae..6205319d0c24 100644
--- a/fs/fs-writeback.c
+++ b/fs/fs-writeback.c
@@ -53,6 +53,7 @@ struct wb_writeback_work {
 	unsigned int for_background:1;
 	unsigned int for_sync:1;	/* sync(2) WB_SYNC_ALL writeback */
 	unsigned int auto_free:1;	/* free on completion */
+	unsigned int start_all:1;	/* nr_pages == 0 (all) writeback */
 	enum wb_reason reason;		/* why was writeback initiated? */
 
 	struct list_head list;		/* pending work list */
@@ -953,12 +954,26 @@ static void wb_start_writeback(struct bdi_writeback *wb, bool range_cyclic,
 		return;
 
 	/*
+	 * All callers of this function want to start writeback of all
+	 * dirty pages. Places like vmscan can call this at a very
+	 * high frequency, causing pointless allocations of tons of
+	 * work items and keeping the flusher threads busy retrieving
+	 * that work. Ensure that we only allow one of them pending and
+	 * inflight at the time
+	 */
+	if (test_bit(WB_start_all, &wb->state))
+		return;
+
+	set_bit(WB_start_all, &wb->state);
+
+	/*
 	 * This is WB_SYNC_NONE writeback, so if allocation fails just
 	 * wakeup the thread for old dirty data writeback
 	 */
 	work = kzalloc(sizeof(*work),
 		       GFP_NOWAIT | __GFP_NOMEMALLOC | __GFP_NOWARN);
 	if (!work) {
+		clear_bit(WB_start_all, &wb->state);
 		trace_writeback_nowork(wb);
 		wb_wakeup(wb);
 		return;
@@ -969,6 +984,7 @@ static void wb_start_writeback(struct bdi_writeback *wb, bool range_cyclic,
 	work->range_cyclic = range_cyclic;
 	work->reason	= reason;
 	work->auto_free	= 1;
+	work->start_all = 1;
 
 	wb_queue_work(wb, work);
 }
@@ -1822,6 +1838,14 @@ static struct wb_writeback_work *get_next_work_item(struct bdi_writeback *wb)
 		list_del_init(&work->list);
 	}
 	spin_unlock_bh(&wb->work_lock);
+
+	/*
+	 * Once we start processing a work item that had !nr_pages,
+	 * clear the wb state bit for that so we can allow more.
+	 */
+	if (work && work->start_all)
+		clear_bit(WB_start_all, &wb->state);
+
 	return work;
 }
 
diff --git a/include/linux/backing-dev-defs.h b/include/linux/backing-dev-defs.h
index 866c433e7d32..420de5c7c7f9 100644
--- a/include/linux/backing-dev-defs.h
+++ b/include/linux/backing-dev-defs.h
@@ -24,6 +24,7 @@ enum wb_state {
 	WB_shutting_down,	/* wb_shutdown() in progress */
 	WB_writeback_running,	/* Writeback is in progress */
 	WB_has_dirty_io,	/* Dirty inodes on ->b_{dirty|io|more_io} */
+	WB_start_all,		/* nr_pages == 0 (all) work pending */
 };
 
 enum wb_congested_state {
-- 
2.7.4

Back to linux.kernel | Previous | NextNext in thread | Find similar | Unroll thread


Thread

[PATCH 7/7] fs-writeback: only allow one inflight and pending full flush Jens Axboe <axboe@kernel.dk> - 2017-09-20 17:40 +0200
  Re: [PATCH 7/7] fs-writeback: only allow one inflight and pending  full flush Christoph Hellwig <hch@infradead.org> - 2017-09-21 17:10 +0200
    Re: [PATCH 7/7] fs-writeback: only allow one inflight and pending  full flush Jens Axboe <axboe@kernel.dk> - 2017-09-21 17:40 +0200
      Re: [PATCH 7/7] fs-writeback: only allow one inflight and pending  full flush Jens Axboe <axboe@kernel.dk> - 2017-09-21 18:10 +0200
        Re: [PATCH 7/7] fs-writeback: only allow one inflight and pending  full flush Christoph Hellwig <hch@infradead.org> - 2017-09-21 19:40 +0200
        Re: [PATCH 7/7] fs-writeback: only allow one inflight and pending  full flush Jan Kara <jack@suse.cz> - 2017-09-25 11:40 +0200
          Re: [PATCH 7/7] fs-writeback: only allow one inflight and pending  full flush Jens Axboe <axboe@kernel.dk> - 2017-09-25 16:50 +0200

csiph-web