linux_old1

History

Michal Hocko 86a294a81f mm, oom, compaction: prevent from should_compact_retry looping for ever for costly orders "mm: consider compaction feedback also for costly allocation" has removed the upper bound for the reclaim/compaction retries based on the number of reclaimed pages for costly orders. While this is desirable the patch did miss a mis interaction between reclaim, compaction and the retry logic. The direct reclaim tries to get zones over min watermark while compaction backs off and returns COMPACT_SKIPPED when all zones are below low watermark + 1<<order gap. If we are getting really close to OOM then __compaction_suitable can keep returning COMPACT_SKIPPED a high order request (e.g. hugetlb order-9) while the reclaim is not able to release enough pages to get us over low watermark. The reclaim is still able to make some progress (usually trashing over few remaining pages) so we are not able to break out from the loop. I have seen this happening with the same test described in "mm: consider compaction feedback also for costly allocation" on a swapless system. The original problem got resolved by "vmscan: consider classzone_idx in compaction_ready" but it shows how things might go wrong when we approach the oom event horizont. The reason why compaction requires being over low rather than min watermark is not clear to me. This check was there essentially since `56de7263fc` ("mm: compaction: direct compact when a high-order allocation fails"). It is clearly an implementation detail though and we shouldn't pull it into the generic retry logic while we should be able to cope with such eventuality. The only place in should_compact_retry where we retry without any upper bound is for compaction_withdrawn() case. Introduce compaction_zonelist_suitable function which checks the given zonelist and returns true only if there is at least one zone which would would unblock __compaction_suitable if more memory got reclaimed. In this implementation it checks __compaction_suitable with NR_FREE_PAGES plus part of the reclaimable memory as the target for the watermark check. The reclaimable memory is reduced linearly by the allocation order. The idea is that we do not want to reclaim all the remaining memory for a single allocation request just unblock __compaction_suitable which doesn't guarantee we will make a further progress. The new helper is then used if compaction_withdrawn() feedback was provided so we do not retry if there is no outlook for a further progress. !costly requests shouldn't be affected much - e.g. order-2 pages would require to have at least 64kB on the reclaimable LRUs while order-9 would need at least 32M which should be enough to not lock up. [vbabka@suse.cz: fix classzone_idx vs. high_zoneidx usage in compaction_zonelist_suitable] [akpm@linux-foundation.org: fix it for Mel's mm-page_alloc-remove-field-from-alloc_context.patch] Signed-off-by: Michal Hocko <mhocko@suse.com> Acked-by: Hillf Danton <hillf.zj@alibaba-inc.com> Acked-by: Vlastimil Babka <vbabka@suse.cz> Cc: David Rientjes <rientjes@google.com> Cc: Johannes Weiner <hannes@cmpxchg.org> Cc: Joonsoo Kim <js1304@gmail.com> Cc: Mel Gorman <mgorman@techsingularity.net> Cc: Tetsuo Handa <penguin-kernel@I-love.SAKURA.ne.jp> Cc: Vladimir Davydov <vdavydov@virtuozzo.com> Signed-off-by: Andrew Morton <akpm@linux-foundation.org> Signed-off-by: Linus Torvalds <torvalds@linux-foundation.org>		2016-05-20 17:58:30 -07:00
..
acpi	Merge branches 'acpi-pci', 'acpi-misc' and 'acpi-tools'	2016-05-16 16:45:48 +02:00
asm-generic	Merge branch 'akpm' (patches from Andrew)	2016-05-19 20:00:06 -07:00
clocksource	clocksource: arm_arch_timer: Remove arch_timer_get_timecounter	2016-05-03 12:54:21 +02:00
crypto	Merge branch 'next' of git://git.kernel.org/pub/scm/linux/kernel/git/jmorris/linux-security	2016-05-19 09:21:36 -07:00
drm	drm: Loongson-3 doesn't fully support wc memory	2016-04-22 10:24:11 +10:00
dt-bindings	- New Drivers	2016-05-20 11:10:24 -07:00
keys	…
kvm	KVM: arm/arm64: vgic: Rely on the GIC driver to parse the firmware tables	2016-05-03 12:54:21 +02:00
linux	mm, oom, compaction: prevent from should_compact_retry looping for ever for costly orders	2016-05-20 17:58:30 -07:00
math-emu	…
media	[media] media: rc: reduce size of struct ir_raw_event	2016-05-07 10:34:17 -03:00
memory	…
misc	cxl: Add kernel API to allow a context to operate with relocate disabled	2016-05-11 21:54:10 +10:00
net	switchdev: pass pointer to fib_info instead of copy	2016-05-17 13:58:49 -04:00
pcmcia	…
ras	…
rdma	IB/security: Restrict use of the write() interface	2016-04-28 12:03:16 -04:00
rxrpc	…
scsi	…
soc	ARC updates for 4.7-rc1	2016-05-19 09:46:18 -07:00
sound	ASoC: Updates for v4.7	2016-05-16 14:59:00 +02:00
target	…
trace	mm, compaction: distinguish between full and partial COMPACT_COMPLETE	2016-05-20 17:58:30 -07:00
uapi	Merge branch 'i2c/for-4.7' of git://git.kernel.org/pub/scm/linux/kernel/git/wsa/linux	2016-05-19 17:48:12 -07:00
video	fbdev: sh_mipi_dsi: remove driver	2016-05-10 11:53:38 +03:00
xen	…
Kbuild	…