Groups | Search | Server Info | Keyboard shortcuts | Login | Register [http] [https] [nntp] [nntps]
Groups > linux.kernel > #1281886
| Path | csiph.com!news.mixmin.net!weretis.net!feeder1.news.weretis.net!news.roellig-ltd.de!open-news-network.org!feeder.erje.net!1.eu.feeder.erje.net!nntpspool01.opticnetworks.net!aioe.org!bofh.it!news.nic.it!robomod |
|---|---|
| From | Michal Hocko <mhocko@kernel.org> |
| Newsgroups | linux.kernel |
| Subject | Re: [PATCH 1/2] mm, oom: Give __GFP_NOFAIL allocations access to memory reserves |
| Date | Wed, 02 Dec 2015 16:10:01 +0100 |
| Message-ID | <qBhcB-nj-17@gated-at.bofh.it> (permalink) |
| References | <qyFO9-56R-3@gated-at.bofh.it> <qyFO9-56R-1@gated-at.bofh.it> <qyFXP-5aJ-9@gated-at.bofh.it> <qyGhb-5xi-7@gated-at.bofh.it> <qyPku-2Rv-3@gated-at.bofh.it> <qz1bY-3hr-11@gated-at.bofh.it> <qAEXF-11R-33@gated-at.bofh.it> |
| X-Original-To | David Rientjes <rientjes@google.com> |
| X-Google-Dkim-Signature | v=1; a=rsa-sha256; c=relaxed/relaxed; d=1e100.net; s=20130820; h=date:from:to:cc:subject:message-id:references:mime-version :content-type:content-disposition:in-reply-to:user-agent; bh=Q/Ryma7a/bpcaqGYCV18vI7Bdt9mfU6WyIeN6gp46x0=; b=SFrqfbNAFuC5pL0BKEU7najASGURzSwZO3yjsbjNaUpkFHNVOyJm5h5n8zu/s1NW7V L3APuOQdI7i+1C8RotZLA6zDh2/lNQXLD5P4+vLFOr31mZJ291ThC1cYB5fsE4VtcEnm z2FDZzgNLt3bTE4gJhFe+77/Q9t5F16YWGckwgbvQV/Dz3wbOVMHF8gM6bBxYgDPCH1/ Mnj/1hJT0oGyVhh+FsYwjs6+CZpGK1LNcXzeVUQ+Tz6NAAIYUUrsUPcm9e1YSVDV8Mp5 sXTThDTSMP3oHA74KhZ4QugPJIvqNo2lISMiy3bFpRmxF+bMC5jm8JnDZetAgM9GCfb0 QxOw== |
| X-Received | by 10.194.175.194 with SMTP id cc2mr5439752wjc.121.1449068852237; Wed, 02 Dec 2015 07:07:32 -0800 (PST) |
| MIME-Version | 1.0 |
| Content-Type | text/plain; charset=us-ascii |
| Content-Disposition | inline |
| User-Agent | Mutt/1.5.24 (2015-08-30) |
| Sender | robomod@news.nic.it |
| List-ID | <linux-kernel.vger.kernel.org> |
| X-Mailing-List | linux-kernel@vger.kernel.org |
| Approved | robomod@news.nic.it |
| Lines | 64 |
| Organization | linux.* mail to news gateway |
| X-Original-Cc | Andrew Morton <akpm@linux-foundation.org>, Mel Gorman <mgorman@suse.de>, Johannes Weiner <hannes@cmpxchg.org>, linux-mm@kvack.org, LKML <linux-kernel@vger.kernel.org> |
| X-Original-Date | Wed, 2 Dec 2015 16:07:30 +0100 |
| X-Original-Message-ID | <20151202150730.GH25284@dhcp22.suse.cz> |
| X-Original-References | <1448448054-804-1-git-send-email-mhocko@kernel.org> <1448448054-804-2-git-send-email-mhocko@kernel.org> <alpine.DEB.2.10.1511250248540.32374@chino.kir.corp.google.com> <20151125111801.GD27283@dhcp22.suse.cz> <alpine.DEB.2.10.1511251254260.24689@chino.kir.corp.google.com> <20151126093427.GA7953@dhcp22.suse.cz> <alpine.DEB.2.10.1511301415010.10460@chino.kir.corp.google.com> |
| X-Original-Sender | linux-kernel-owner@vger.kernel.org |
| Xref | csiph.com linux.kernel:1281886 |
Show key headers only | View raw
On Mon 30-11-15 14:17:03, David Rientjes wrote:
> On Thu, 26 Nov 2015, Michal Hocko wrote:
>
> > > > diff --git a/mm/page_alloc.c b/mm/page_alloc.c
> > > > index 8034909faad2..94b04c1e894a 100644
> > > > --- a/mm/page_alloc.c
> > > > +++ b/mm/page_alloc.c
> > > > @@ -2766,8 +2766,13 @@ __alloc_pages_may_oom(gfp_t gfp_mask, unsigned int order,
> > > > goto out;
> > > > }
> > > > /* Exhausted what can be done so it's blamo time */
> > > > - if (out_of_memory(&oc) || WARN_ON_ONCE(gfp_mask & __GFP_NOFAIL))
> > > > + if (out_of_memory(&oc) || WARN_ON_ONCE(gfp_mask & __GFP_NOFAIL)) {
> > > > *did_some_progress = 1;
> > > > +
> > > > + if (gfp_mask & __GFP_NOFAIL)
> > > > + page = get_page_from_freelist(gfp_mask, order,
> > > > + ALLOC_NO_WATERMARKS, ac);
> > > > + }
> > > > out:
> > > > mutex_unlock(&oom_lock);
> > > > return page;
> > >
> > > Well, sure, that's one way to do it, but for cpuset users, wouldn't this
> > > lead to a depletion of the first system zone since you've dropped
> > > ALLOC_CPUSET and are doing ALLOC_NO_WATERMARKS in the same call?
> >
> > Are you suggesting to do?
> > if (gfp_mask & __GFP_NOFAIL) {
> > page = get_page_from_freelist(gfp_mask, order,
> > ALLOC_NO_WATERMARKS|ALLOC_CPUSET, ac);
> > /*
> > * fallback to ignore cpuset if our nodes are
> > * depleted
> > */
> > if (!page)
> > get_page_from_freelist(gfp_mask, order,
> > ALLOC_NO_WATERMARKS, ac);
> > }
> >
> > I am not really sure this worth complication.
>
> I'm objecting to the ability of a process that is doing a __GFP_NOFAIL
> allocation, which has been disallowed access from allocating on certain
> mems through cpusets, to cause an oom condition on those disallowed nodes,
> yes.
That ability will be there even with the fallback mechanism. My primary
objections was that the fallback is unnecessarily complex without any
evidence that such a situation would happen in the real life often
enought to bother about it. __GFP_NOFAIL allocations are and should be
rare and any runaway triggerable from the userspace is a kernel bug.
Anyway, as you seem to feel really strongly about this I will post v2
with the above fallback. This is a superslow path anyway...
--
Michal Hocko
SUSE Labs
--
To unsubscribe from this list: send the line "unsubscribe linux-kernel" in
the body of a message to majordomo@vger.kernel.org
More majordomo info at http://vger.kernel.org/majordomo-info.html
Please read the FAQ at http://www.tux.org/lkml/
Back to linux.kernel | Previous | Next — Previous in thread | Find similar | Unroll thread
Re: [PATCH 1/2] mm, oom: Give __GFP_NOFAIL allocations access to memory reserves David Rientjes <rientjes@google.com> - 2015-11-30 23:20 +0100 Re: [PATCH 1/2] mm, oom: Give __GFP_NOFAIL allocations access to memory reserves Michal Hocko <mhocko@kernel.org> - 2015-12-02 16:10 +0100
csiph-web