Groups | Search | Server Info | Keyboard shortcuts | Login | Register [http] [https] [nntp] [nntps]


Groups > linux.kernel > #1281886

Re: [PATCH 1/2] mm, oom: Give __GFP_NOFAIL allocations access to memory reserves

Path csiph.com!news.mixmin.net!weretis.net!feeder1.news.weretis.net!news.roellig-ltd.de!open-news-network.org!feeder.erje.net!1.eu.feeder.erje.net!nntpspool01.opticnetworks.net!aioe.org!bofh.it!news.nic.it!robomod
From Michal Hocko <mhocko@kernel.org>
Newsgroups linux.kernel
Subject Re: [PATCH 1/2] mm, oom: Give __GFP_NOFAIL allocations access to memory reserves
Date Wed, 02 Dec 2015 16:10:01 +0100
Message-ID <qBhcB-nj-17@gated-at.bofh.it> (permalink)
References <qyFO9-56R-3@gated-at.bofh.it> <qyFO9-56R-1@gated-at.bofh.it> <qyFXP-5aJ-9@gated-at.bofh.it> <qyGhb-5xi-7@gated-at.bofh.it> <qyPku-2Rv-3@gated-at.bofh.it> <qz1bY-3hr-11@gated-at.bofh.it> <qAEXF-11R-33@gated-at.bofh.it>
X-Original-To David Rientjes <rientjes@google.com>
X-Google-Dkim-Signature v=1; a=rsa-sha256; c=relaxed/relaxed; d=1e100.net; s=20130820; h=date:from:to:cc:subject:message-id:references:mime-version :content-type:content-disposition:in-reply-to:user-agent; bh=Q/Ryma7a/bpcaqGYCV18vI7Bdt9mfU6WyIeN6gp46x0=; b=SFrqfbNAFuC5pL0BKEU7najASGURzSwZO3yjsbjNaUpkFHNVOyJm5h5n8zu/s1NW7V L3APuOQdI7i+1C8RotZLA6zDh2/lNQXLD5P4+vLFOr31mZJ291ThC1cYB5fsE4VtcEnm z2FDZzgNLt3bTE4gJhFe+77/Q9t5F16YWGckwgbvQV/Dz3wbOVMHF8gM6bBxYgDPCH1/ Mnj/1hJT0oGyVhh+FsYwjs6+CZpGK1LNcXzeVUQ+Tz6NAAIYUUrsUPcm9e1YSVDV8Mp5 sXTThDTSMP3oHA74KhZ4QugPJIvqNo2lISMiy3bFpRmxF+bMC5jm8JnDZetAgM9GCfb0 QxOw==
X-Received by 10.194.175.194 with SMTP id cc2mr5439752wjc.121.1449068852237; Wed, 02 Dec 2015 07:07:32 -0800 (PST)
MIME-Version 1.0
Content-Type text/plain; charset=us-ascii
Content-Disposition inline
User-Agent Mutt/1.5.24 (2015-08-30)
Sender robomod@news.nic.it
List-ID <linux-kernel.vger.kernel.org>
X-Mailing-List linux-kernel@vger.kernel.org
Approved robomod@news.nic.it
Lines 64
Organization linux.* mail to news gateway
X-Original-Cc Andrew Morton <akpm@linux-foundation.org>, Mel Gorman <mgorman@suse.de>, Johannes Weiner <hannes@cmpxchg.org>, linux-mm@kvack.org, LKML <linux-kernel@vger.kernel.org>
X-Original-Date Wed, 2 Dec 2015 16:07:30 +0100
X-Original-Message-ID <20151202150730.GH25284@dhcp22.suse.cz>
X-Original-References <1448448054-804-1-git-send-email-mhocko@kernel.org> <1448448054-804-2-git-send-email-mhocko@kernel.org> <alpine.DEB.2.10.1511250248540.32374@chino.kir.corp.google.com> <20151125111801.GD27283@dhcp22.suse.cz> <alpine.DEB.2.10.1511251254260.24689@chino.kir.corp.google.com> <20151126093427.GA7953@dhcp22.suse.cz> <alpine.DEB.2.10.1511301415010.10460@chino.kir.corp.google.com>
X-Original-Sender linux-kernel-owner@vger.kernel.org
Xref csiph.com linux.kernel:1281886

Show key headers only | View raw


On Mon 30-11-15 14:17:03, David Rientjes wrote:
> On Thu, 26 Nov 2015, Michal Hocko wrote:
> 
> > > > diff --git a/mm/page_alloc.c b/mm/page_alloc.c
> > > > index 8034909faad2..94b04c1e894a 100644
> > > > --- a/mm/page_alloc.c
> > > > +++ b/mm/page_alloc.c
> > > > @@ -2766,8 +2766,13 @@ __alloc_pages_may_oom(gfp_t gfp_mask, unsigned int order,
> > > >  			goto out;
> > > >  	}
> > > >  	/* Exhausted what can be done so it's blamo time */
> > > > -	if (out_of_memory(&oc) || WARN_ON_ONCE(gfp_mask & __GFP_NOFAIL))
> > > > +	if (out_of_memory(&oc) || WARN_ON_ONCE(gfp_mask & __GFP_NOFAIL)) {
> > > >  		*did_some_progress = 1;
> > > > +
> > > > +		if (gfp_mask & __GFP_NOFAIL)
> > > > +			page = get_page_from_freelist(gfp_mask, order,
> > > > +					ALLOC_NO_WATERMARKS, ac);
> > > > +	}
> > > >  out:
> > > >  	mutex_unlock(&oom_lock);
> > > >  	return page;
> > > 
> > > Well, sure, that's one way to do it, but for cpuset users, wouldn't this 
> > > lead to a depletion of the first system zone since you've dropped 
> > > ALLOC_CPUSET and are doing ALLOC_NO_WATERMARKS in the same call?  
> > 
> > Are you suggesting to do?
> > 		if (gfp_mask & __GFP_NOFAIL) {
> > 			page = get_page_from_freelist(gfp_mask, order,
> > 					ALLOC_NO_WATERMARKS|ALLOC_CPUSET, ac);
> > 			/*
> > 			 * fallback to ignore cpuset if our nodes are
> > 			 * depleted
> > 			 */
> > 			if (!page)
> > 				get_page_from_freelist(gfp_mask, order,
> > 					ALLOC_NO_WATERMARKS, ac);
> > 		}
> > 
> > I am not really sure this worth complication.
> 
> I'm objecting to the ability of a process that is doing a __GFP_NOFAIL 
> allocation, which has been disallowed access from allocating on certain 
> mems through cpusets, to cause an oom condition on those disallowed nodes, 
> yes.

That ability will be there even with the fallback mechanism. My primary
objections was that the fallback is unnecessarily complex without any
evidence that such a situation would happen in the real life often
enought to bother about it. __GFP_NOFAIL allocations are and should be
rare and any runaway triggerable from the userspace is a kernel bug.

Anyway, as you seem to feel really strongly about this I will post v2
with the above fallback. This is a superslow path anyway...

-- 
Michal Hocko
SUSE Labs
--
To unsubscribe from this list: send the line "unsubscribe linux-kernel" in
the body of a message to majordomo@vger.kernel.org
More majordomo info at  http://vger.kernel.org/majordomo-info.html
Please read the FAQ at  http://www.tux.org/lkml/

Back to linux.kernel | Previous | Next — Previous in thread | Find similar | Unroll thread


Thread

Re: [PATCH 1/2] mm, oom: Give __GFP_NOFAIL allocations access to  memory reserves David Rientjes <rientjes@google.com> - 2015-11-30 23:20 +0100
  Re: [PATCH 1/2] mm, oom: Give __GFP_NOFAIL allocations access to  memory reserves Michal Hocko <mhocko@kernel.org> - 2015-12-02 16:10 +0100

csiph-web