Re: [PATCH] RDS: sync congestion map updating
From: santosh.shilimkar-QHcLZuEGTsvQT0dZR+AlfA@public.gmane.org <hidden>
Date: 2016-04-02 04:30:54
Also in:
linux-rdma
On 4/1/16 6:14 PM, Leon Romanovsky wrote:
On Fri, Apr 01, 2016 at 12:47:24PM -0700, santosh shilimkar wrote:quoted
(cc-ing netdev) On 3/30/2016 7:59 PM, Wengang Wang wrote:quoted
在 2016年03月31日 09:51, Wengang Wang 写道:quoted
在 2016年03月31日 01:16, santosh shilimkar 写道:quoted
Hi Wengang, On 3/30/2016 9:19 AM, Leon Romanovsky wrote:quoted
On Wed, Mar 30, 2016 at 05:08:22PM +0800, Wengang Wang wrote:quoted
Problem is found that some among a lot of parallel RDS communications hang. In my test ten or so among 33 communications hang. The send requests got -ENOBUF error meaning the peer socket (port) is congested. But meanwhile, peer socket (port) is not congested. The congestion map updating can happen in two paths: one is in rds_recvmsg path and the other is when it receives packets from the hardware. There is no synchronization when updating the congestion map. So a bit operation (clearing) in the rds_recvmsg path can be skipped by another bit operation (setting) in hardware packet receving path.To be more detailed. Here, the two paths (user calls recvmsg and hardware receives data) are for different rds socks. thus the rds_sock->rs_recv_lock is not helpful to sync the updating on congestion map.For archive purpose, let me try to conclude the thread. I synced with Wengang offlist and came up with below fix. I was under impression that __set_bit_le() was atmoic version. After fixing it like patch(end of the email), the bug gets addressed. I will probably send this as fix for stable as well. From 5614b61f6fdcd6ae0c04e50b97efd13201762294 Mon Sep 17 00:00:00 2001 From: Santosh Shilimkar <redacted> Date: Wed, 30 Mar 2016 23:26:47 -0700 Subject: [PATCH] RDS: Fix the atomicity for congestion map update Two different threads with different rds sockets may be in rds_recv_rcvbuf_delta() via receive path. If their ports both map to the same word in the congestion map, then using non-atomic ops to update it could cause the map to be incorrect. Lets use atomics to avoid such an issue. Full credit to Wengang [off-list ref] for finding the issue, analysing it and also pointing out to offending code with spin lock based fix.I'm glad that you solved the issue without spinlocks. Out of curiosity, I see that this patch is needed to be sent to Dave and applied by him. Is it right?
Right. I was planning send this one along with one more fix together on netdev for Dave to pick it up.
➜ linus-tree git:(master) ./scripts/get_maintainer.pl -f net/rds/cong.c Santosh Shilimkar [off-list ref] (supporter:RDS - RELIABLE DATAGRAM SOCKETS) "David S. Miller" [off-list ref] (maintainer:NETWORKING [GENERAL]) netdev-u79uwXL29TY76Z2rM5mHXA@public.gmane.org (open list:RDS - RELIABLE DATAGRAM SOCKETS) linux-rdma-u79uwXL29TY76Z2rM5mHXA@public.gmane.org (open list:RDS - RELIABLE DATAGRAM SOCKETS) rds-devel-N0ozoZBvEnrZJqsBc5GL+g@public.gmane.org (moderated list:RDS - RELIABLE DATAGRAM SOCKETS) linux-kernel-u79uwXL29TY76Z2rM5mHXA@public.gmane.org (open list)quoted
Signed-off-by: Wengang Wang <redacted> Signed-off-by: Santosh Shilimkar <redacted>Reviewed-by: Leon Romanovsky <leon-2ukJVAZIZ/Y@public.gmane.org>
Thanks for review. -- To unsubscribe from this list: send the line "unsubscribe linux-rdma" in the body of a message to majordomo-u79uwXL29TY76Z2rM5mHXA@public.gmane.org More majordomo info at http://vger.kernel.org/majordomo-info.html