linux-sg2042

History

Xue jiufei b246d3d11e ocfs2: fix a deadlock while o2net_wq doing direct memory reclaim Fix a deadlock problem caused by direct memory reclaim in o2net_wq. The situation is as follows: 1) Receive a connect message from another node, node queues a work_struct o2net_listen_work. 2) o2net_wq processes this work and call the following functions: o2net_wq -> o2net_accept_one -> sock_create_lite -> sock_alloc() -> kmem_cache_alloc with GFP_KERNEL -> ____cache_alloc_node ->__alloc_pages_nodemask -> do_try_to_free_pages -> shrink_slab -> evict -> ocfs2_evict_inode -> ocfs2_drop_lock -> dlmunlock -> o2net_send_message_vec then o2net_wq wait for the unlock reply from master. 3) tcp layer received the reply, call o2net_data_ready() and queue sc_rx_work, waiting o2net_wq to process this work. 4) o2net_wq is a single thread workqueue, it process the work one by one. Right now it is still doing o2net_listen_work and cannot handle sc_rx_work. so we deadlock. Junxiao Bi's patch "mm: clear __GFP_FS when PF_MEMALLOC_NOIO is set" (http://ozlabs.org/~akpm/mmots/broken-out/mm-clear-__gfp_fs-when-pf_memalloc_noio-is-set.patch) clears __GFP_FS in memalloc_noio_flags() besides __GFP_IO. We use memalloc_noio_save() to set process flag PF_MEMALLOC_NOIO so that all allocations done by this process are done as if GFP_NOIO was specified. We are not reentering filesystem while doing memory reclaim. Signed-off-by: joyce.xue <xuejiufei@huawei.com> Cc: Junxiao Bi <junxiao.bi@oracle.com> Cc: Joel Becker <jlbec@evilplan.org> Cc: Mark Fasheh <mfasheh@suse.com> Signed-off-by: Andrew Morton <akpm@linux-foundation.org> Signed-off-by: Linus Torvalds <torvalds@linux-foundation.org>		2014-10-09 22:25:58 -04:00
..
Makefile	ocfs2: remove versioning information	2014-01-21 16:19:41 -08:00
heartbeat.c	ocfs2: fix deadlock between o2hb thread and o2net_wq	2014-10-09 22:25:47 -04:00
heartbeat.h	ocfs2: fix deadlock between o2hb thread and o2net_wq	2014-10-09 22:25:47 -04:00
masklog.c	ocfs2: Remove masklog ML_UPTODATE.	2011-02-24 16:22:20 +08:00
masklog.h	ocfs2: don't spam on -EDQUOT	2013-11-13 12:09:01 +09:00
netdebug.c	fs/ocfs2/cluster/netdebug.c: use seq_open_private() not seq_open()	2014-10-09 22:25:47 -04:00
nodemanager.c	ocfs2: remove versioning information	2014-01-21 16:19:41 -08:00
nodemanager.h	ocfs2/cluster: Make fence method configurable - v2	2009-12-02 16:49:26 -08:00
ocfs2_heartbeat.h	ocfs2: warn the user on a dead timeout mismatch	2006-06-29 15:45:35 -07:00
ocfs2_nodemanager.h	ocfs2/dlm: Add message DLM_QUERY_REGION	2010-10-09 10:26:23 -07:00
quorum.c	ocfs2: quorum: add a log for node not fenced	2014-08-29 16:28:17 -07:00
quorum.h	…
sys.c	VERIFY_OCTAL_PERMISSIONS: stricter checking for sysfs perms.	2014-03-24 12:21:00 +10:30
sys.h	…
tcp.c	ocfs2: fix a deadlock while o2net_wq doing direct memory reclaim	2014-10-09 22:25:58 -04:00
tcp.h	ocfs2: o2net: set tcp user timeout to max value	2014-08-29 16:28:16 -07:00
tcp_internal.h	net: Fix use after free by removing length arg from sk_data_ready callbacks.	2014-04-11 16:15:36 -04:00