[PATCH v5 04/45] percpu_rwlock: Implement the core design of Per-CPU... | linux-arm-kernel

[PATCH v5 00/45] CPU hotplug: stop_machine()-free CPU hotplug · Srivatsa S. Bhat <hidden> · 2013-01-22
[PATCH v5 01/45] percpu_rwlock: Introduce the global reader-writer lock backend · Srivatsa S. Bhat <hidden> · 2013-01-22
Re: [PATCH v5 01/45] percpu_rwlock: Introduce the global reader-writer lock backend · Stephen Hemminger <stephen@networkplumber.org> · 2013-01-22
Re: [PATCH v5 01/45] percpu_rwlock: Introduce the global reader-writer lock backend · Srivatsa S. Bhat <hidden> · 2013-01-22
Re: [PATCH v5 01/45] percpu_rwlock: Introduce the global reader-writer lock backend · Steven Rostedt <rostedt@goodmis.org> · 2013-01-22
Re: [PATCH v5 01/45] percpu_rwlock: Introduce the global reader-writer lock backend · Srivatsa S. Bhat <hidden> · 2013-01-22
Re: [PATCH v5 01/45] percpu_rwlock: Introduce the global reader-writer lock backend · Steven Rostedt <rostedt@goodmis.org> · 2013-01-22
Re: [PATCH v5 01/45] percpu_rwlock: Introduce the global reader-writer lock backend · Michel Lespinasse <hidden> · 2013-01-24
[PATCH v5 01/45] percpu_rwlock: Introduce the global reader-writer lock backend · oleg@redhat.com (Oleg Nesterov) · 2013-01-24
Re: [PATCH v5 01/45] percpu_rwlock: Introduce the global reader-writer lock backend · David Howells <dhowells@redhat.com> · 2013-02-11
Re: [PATCH v5 01/45] percpu_rwlock: Introduce the global reader-writer lock backend · Srivatsa S. Bhat <hidden> · 2013-02-11
[PATCH v5 02/45] percpu_rwlock: Introduce per-CPU variables for the reader and the writer · Srivatsa S. Bhat <hidden> · 2013-01-22
[PATCH v5 03/45] percpu_rwlock: Provide a way to define and init percpu-rwlocks at compile time · Srivatsa S. Bhat <hidden> · 2013-01-22
[PATCH v5 04/45] percpu_rwlock: Implement the core design of Per-CPU Reader-Writer Locks · Srivatsa S. Bhat <hidden> · 2013-01-22
Re: [PATCH v5 04/45] percpu_rwlock: Implement the core design of Per-CPU Reader-Writer Locks · Tejun Heo <tj@kernel.org> · 2013-01-23
Re: [PATCH v5 04/45] percpu_rwlock: Implement the core design of Per-CPU Reader-Writer Locks · Srivatsa S. Bhat <hidden> · 2013-01-23
Re: [PATCH v5 04/45] percpu_rwlock: Implement the core design of Per-CPU Reader-Writer Locks · Tejun Heo <tj@kernel.org> · 2013-01-23
Re: [PATCH v5 04/45] percpu_rwlock: Implement the core design of Per-CPU Reader-Writer Locks · Srivatsa S. Bhat <hidden> · 2013-01-24
Re: [PATCH v5 04/45] percpu_rwlock: Implement the core design of Per-CPU Reader-Writer Locks · Paul E. McKenney <hidden> · 2013-02-08
Re: [PATCH v5 04/45] percpu_rwlock: Implement the core design of Per-CPU Reader-Writer Locks · Oleg Nesterov <oleg@redhat.com> · 2013-02-10
Re: [PATCH v5 04/45] percpu_rwlock: Implement the core design of Per-CPU Reader-Writer Locks · Srivatsa S. Bhat <hidden> · 2013-02-10
Re: [PATCH v5 04/45] percpu_rwlock: Implement the core design of Per-CPU Reader-Writer Locks · Oleg Nesterov <oleg@redhat.com> · 2013-02-10
Re: [PATCH v5 04/45] percpu_rwlock: Implement the core design of Per-CPU Reader-Writer Locks · Srivatsa S. Bhat <hidden> · 2013-02-10
Re: [PATCH v5 04/45] percpu_rwlock: Implement the core design of Per-CPU Reader-Writer Locks · Paul E. McKenney <hidden> · 2013-02-10
Re: [PATCH v5 04/45] percpu_rwlock: Implement the core design of Per-CPU Reader-Writer Locks · Paul E. McKenney <hidden> · 2013-02-10
Re: [PATCH v5 04/45] percpu_rwlock: Implement the core design of Per-CPU Reader-Writer Locks · Paul E. McKenney <hidden> · 2013-02-12
Re: [PATCH v5 04/45] percpu_rwlock: Implement the core design of Per-CPU Reader-Writer Locks · Srivatsa S. Bhat <hidden> · 2013-02-10
Re: [PATCH v5 04/45] percpu_rwlock: Implement the core design of Per-CPU Reader-Writer Locks · Paul E. McKenney <hidden> · 2013-02-10
Re: [PATCH v5 04/45] percpu_rwlock: Implement the core design of Per-CPU Reader-Writer Locks · Srivatsa S. Bhat <hidden> · 2013-02-10
Re: [PATCH v5 04/45] percpu_rwlock: Implement the core design of Per-CPU Reader-Writer Locks · Oleg Nesterov <oleg@redhat.com> · 2013-02-10
Re: [PATCH v5 04/45] percpu_rwlock: Implement the core design of Per-CPU Reader-Writer Locks · Srivatsa S. Bhat <hidden> · 2013-02-10
[PATCH v5 05/45] percpu_rwlock: Make percpu-rwlocks IRQ-safe, optimally · Srivatsa S. Bhat <hidden> · 2013-01-22
Re: [PATCH v5 05/45] percpu_rwlock: Make percpu-rwlocks IRQ-safe, optimally · Paul E. McKenney <hidden> · 2013-02-08
Re: [PATCH v5 05/45] percpu_rwlock: Make percpu-rwlocks IRQ-safe, optimally · Srivatsa S. Bhat <hidden> · 2013-02-10
Re: [PATCH v5 05/45] percpu_rwlock: Make percpu-rwlocks IRQ-safe, optimally · Oleg Nesterov <oleg@redhat.com> · 2013-02-10
Re: [PATCH v5 05/45] percpu_rwlock: Make percpu-rwlocks IRQ-safe, optimally · Srivatsa S. Bhat <hidden> · 2013-02-10
[PATCH v5 06/45] percpu_rwlock: Allow writers to be readers, and add lockdep annotations · Srivatsa S. Bhat <hidden> · 2013-01-22
Re: [PATCH v5 06/45] percpu_rwlock: Allow writers to be readers, and add lockdep annotations · Paul E. McKenney <hidden> · 2013-02-08
Re: [PATCH v5 06/45] percpu_rwlock: Allow writers to be readers, and add lockdep annotations · Srivatsa S. Bhat <hidden> · 2013-02-10
[PATCH v5 07/45] CPU hotplug: Provide APIs to prevent CPU offline from atomic context · Srivatsa S. Bhat <hidden> · 2013-01-22
Re: [PATCH v5 07/45] CPU hotplug: Provide APIs to prevent CPU offline from atomic context · Paul E. McKenney <hidden> · 2013-02-08
[PATCH v5 08/45] CPU hotplug: Convert preprocessor macros to static inline functions · Srivatsa S. Bhat <hidden> · 2013-01-22
Re: [PATCH v5 08/45] CPU hotplug: Convert preprocessor macros to static inline functions · Paul E. McKenney <hidden> · 2013-02-08
[PATCH v5 09/45] smp, cpu hotplug: Fix smp_call_function_*() to prevent CPU offline properly · Srivatsa S. Bhat <hidden> · 2013-01-22
Re: [PATCH v5 09/45] smp, cpu hotplug: Fix smp_call_function_*() to prevent CPU offline properly · Paul E. McKenney <hidden> · 2013-02-09
Re: [PATCH v5 09/45] smp, cpu hotplug: Fix smp_call_function_*() to prevent CPU offline properly · Srivatsa S. Bhat <hidden> · 2013-02-10
Re: [PATCH v5 09/45] smp, cpu hotplug: Fix smp_call_function_*() to prevent CPU offline properly · Paul E. McKenney <hidden> · 2013-02-10
Re: [PATCH v5 09/45] smp, cpu hotplug: Fix smp_call_function_*() to prevent CPU offline properly · Srivatsa S. Bhat <hidden> · 2013-02-10
[PATCH v5 10/45] smp, cpu hotplug: Fix on_each_cpu_*() to prevent CPU offline properly · Srivatsa S. Bhat <hidden> · 2013-01-22
[PATCH v5 11/45] sched/timer: Use get/put_online_cpus_atomic() to prevent CPU offline · Srivatsa S. Bhat <hidden> · 2013-01-22
[PATCH v5 12/45] sched/migration: Use raw_spin_lock/unlock since interrupts are already disabled · Srivatsa S. Bhat <hidden> · 2013-01-22
[PATCH v5 13/45] sched/rt: Use get/put_online_cpus_atomic() to prevent CPU offline · Srivatsa S. Bhat <hidden> · 2013-01-22
[PATCH v5 14/45] rcu, CPU hotplug: Fix comment referring to stop_machine() · Srivatsa S. Bhat <hidden> · 2013-01-22
Re: [PATCH v5 14/45] rcu, CPU hotplug: Fix comment referring to stop_machine() · Paul E. McKenney <hidden> · 2013-02-09
Re: [PATCH v5 14/45] rcu, CPU hotplug: Fix comment referring to stop_machine() · Srivatsa S. Bhat <hidden> · 2013-02-10
[PATCH v5 15/45] tick: Use get/put_online_cpus_atomic() to prevent CPU offline · Srivatsa S. Bhat <hidden> · 2013-01-22
[PATCH v5 16/45] time/clocksource: Use get/put_online_cpus_atomic() to prevent CPU offline · Srivatsa S. Bhat <hidden> · 2013-01-22
[PATCH v5 17/45] softirq: Use get/put_online_cpus_atomic() to prevent CPU offline · Srivatsa S. Bhat <hidden> · 2013-01-22
[PATCH v5 18/45] irq: Use get/put_online_cpus_atomic() to prevent CPU offline · Srivatsa S. Bhat <hidden> · 2013-01-22
[PATCH v5 19/45] net: Use get/put_online_cpus_atomic() to prevent CPU offline · Srivatsa S. Bhat <hidden> · 2013-01-22
[PATCH v5 20/45] block: Use get/put_online_cpus_atomic() to prevent CPU offline · Srivatsa S. Bhat <hidden> · 2013-01-22
[PATCH v5 21/45] crypto: pcrypt - Protect access to cpu_online_mask with get/put_online_cpus() · Srivatsa S. Bhat <hidden> · 2013-01-22
[PATCH v5 22/45] infiniband: ehca: Use get/put_online_cpus_atomic() to prevent CPU offline · Srivatsa S. Bhat <hidden> · 2013-01-22
[PATCH v5 23/45] [SCSI] fcoe: Use get/put_online_cpus_atomic() to prevent CPU offline · Srivatsa S. Bhat <hidden> · 2013-01-22
[PATCH v5 24/45] staging: octeon: Use get/put_online_cpus_atomic() to prevent CPU offline · Srivatsa S. Bhat <hidden> · 2013-01-22
[PATCH v5 25/45] x86: Use get/put_online_cpus_atomic() to prevent CPU offline · Srivatsa S. Bhat <hidden> · 2013-01-22
[PATCH v5 27/45] KVM: Use get/put_online_cpus_atomic() to prevent CPU offline from atomic context · Srivatsa S. Bhat <hidden> · 2013-01-22
[PATCH v5 28/45] kvm/vmx: Use get/put_online_cpus_atomic() to prevent CPU offline · Srivatsa S. Bhat <hidden> · 2013-01-22
[PATCH v5 26/45] perf/x86: Use get/put_online_cpus_atomic() to prevent CPU offline · Srivatsa S. Bhat <hidden> · 2013-01-22
[PATCH v5 29/45] x86/xen: Use get/put_online_cpus_atomic() to prevent CPU offline · Srivatsa S. Bhat <hidden> · 2013-01-22
Re: [PATCH v5 29/45] x86/xen: Use get/put_online_cpus_atomic() to prevent CPU offline · Konrad Rzeszutek Wilk <konrad@kernel.org> · 2013-02-19
Re: [PATCH v5 29/45] x86/xen: Use get/put_online_cpus_atomic() to prevent CPU offline · Srivatsa S. Bhat <hidden> · 2013-02-19
[PATCH v5 30/45] alpha/smp: Use get/put_online_cpus_atomic() to prevent CPU offline · Srivatsa S. Bhat <hidden> · 2013-01-22
[PATCH v5 31/45] blackfin/smp: Use get/put_online_cpus_atomic() to prevent CPU offline · Srivatsa S. Bhat <hidden> · 2013-01-22
Re: [PATCH v5 31/45] blackfin/smp: Use get/put_online_cpus_atomic() to prevent CPU offline · Bob Liu <hidden> · 2013-01-28
Re: [PATCH v5 31/45] blackfin/smp: Use get/put_online_cpus_atomic() to prevent CPU offline · Tejun Heo <tj@kernel.org> · 2013-01-28
Re: [PATCH v5 31/45] blackfin/smp: Use get/put_online_cpus_atomic() to prevent CPU offline · Srivatsa S. Bhat <hidden> · 2013-01-29
[PATCH v5 32/45] cris/smp: Use get/put_online_cpus_atomic() to prevent CPU offline · Srivatsa S. Bhat <hidden> · 2013-01-22
[PATCH v5 33/45] hexagon/smp: Use get/put_online_cpus_atomic() to prevent CPU offline · Srivatsa S. Bhat <hidden> · 2013-01-22
[PATCH v5 34/45] ia64: Use get/put_online_cpus_atomic() to prevent CPU offline · Srivatsa S. Bhat <hidden> · 2013-01-22
[PATCH v5 35/45] m32r: Use get/put_online_cpus_atomic() to prevent CPU offline · Srivatsa S. Bhat <hidden> · 2013-01-22
[PATCH v5 36/45] MIPS: Use get/put_online_cpus_atomic() to prevent CPU offline · Srivatsa S. Bhat <hidden> · 2013-01-22
[PATCH v5 37/45] mn10300: Use get/put_online_cpus_atomic() to prevent CPU offline · Srivatsa S. Bhat <hidden> · 2013-01-22
[PATCH v5 38/45] parisc: Use get/put_online_cpus_atomic() to prevent CPU offline · Srivatsa S. Bhat <hidden> · 2013-01-22
[PATCH v5 39/45] powerpc: Use get/put_online_cpus_atomic() to prevent CPU offline · Srivatsa S. Bhat <hidden> · 2013-01-22
[PATCH v5 40/45] sh: Use get/put_online_cpus_atomic() to prevent CPU offline · Srivatsa S. Bhat <hidden> · 2013-01-22
[PATCH v5 41/45] sparc: Use get/put_online_cpus_atomic() to prevent CPU offline · Srivatsa S. Bhat <hidden> · 2013-01-22
[PATCH v5 42/45] tile: Use get/put_online_cpus_atomic() to prevent CPU offline · Srivatsa S. Bhat <hidden> · 2013-01-22
[PATCH v5 43/45] cpu: No more __stop_machine() in _cpu_down() · Srivatsa S. Bhat <hidden> · 2013-01-22
[PATCH v5 44/45] CPU hotplug, stop_machine: Decouple CPU hotplug from stop_machine() in Kconfig · Srivatsa S. Bhat <hidden> · 2013-01-22
Re: [PATCH v5 44/45] CPU hotplug, stop_machine: Decouple CPU hotplug from stop_machine() in Kconfig · Paul E. McKenney <hidden> · 2013-02-09
Re: [PATCH v5 44/45] CPU hotplug, stop_machine: Decouple CPU hotplug from stop_machine() in Kconfig · Srivatsa S. Bhat <hidden> · 2013-02-10
[PATCH v5 45/45] Documentation/cpu-hotplug: Remove references to stop_machine() · Srivatsa S. Bhat <hidden> · 2013-01-22
Re: [PATCH v5 45/45] Documentation/cpu-hotplug: Remove references to stop_machine() · Paul E. McKenney <hidden> · 2013-02-09
Re: [PATCH v5 00/45] CPU hotplug: stop_machine()-free CPU hotplug · Srivatsa S. Bhat <hidden> · 2013-02-04
Re: [PATCH v5 00/45] CPU hotplug: stop_machine()-free CPU hotplug · Rusty Russell <hidden> · 2013-02-07
Re: [PATCH v5 00/45] CPU hotplug: stop_machine()-free CPU hotplug · Srivatsa S. Bhat <hidden> · 2013-02-07
Re: [PATCH v5 00/45] CPU hotplug: stop_machine()-free CPU hotplug · Russell King - ARM Linux <hidden> · 2013-02-08
Re: [PATCH v5 00/45] CPU hotplug: stop_machine()-free CPU hotplug · Srivatsa S. Bhat <hidden> · 2013-02-08
Re: [PATCH v5 00/45] CPU hotplug: stop_machine()-free CPU hotplug · Srivatsa S. Bhat <hidden> · 2013-02-08
Re: [PATCH v5 00/45] CPU hotplug: stop_machine()-free CPU hotplug · Vincent Guittot <vincent.guittot@linaro.org> · 2013-02-11
Re: [PATCH v5 00/45] CPU hotplug: stop_machine()-free CPU hotplug · Srivatsa S. Bhat <hidden> · 2013-02-11
Re: [PATCH v5 00/45] CPU hotplug: stop_machine()-free CPU hotplug · Paul E. McKenney <hidden> · 2013-02-11
Re: [PATCH v5 00/45] CPU hotplug: stop_machine()-free CPU hotplug · Srivatsa S. Bhat <hidden> · 2013-02-12

[PATCH v5 04/45] percpu_rwlock: Implement the core design of Per-CPU Reader-Writer Locks

From: oleg@redhat.com (Oleg Nesterov)
Date: 2013-02-10 18:09:44
Also in: linux-arch, linux-pm, linuxppc-dev, lkml, netdev

On 02/08, Paul E. McKenney wrote:

On Tue, Jan 22, 2013 at 01:03:53PM +0530, Srivatsa S. Bhat wrote:

quoted

 void percpu_read_unlock(struct percpu_rwlock *pcpu_rwlock)
 {
-	read_unlock(&pcpu_rwlock->global_rwlock);

We need an smp_mb() here to keep the critical section ordered before the
this_cpu_dec() below.  Otherwise, if a writer shows up just after we
exit the fastpath, that writer is not guaranteed to see the effects of
our critical section.  Equivalently, the prior read-side critical section
just might see some of the writer's updates, which could be a bit of
a surprise to the reader.

Agreed, we should not assume that a "reader" doesn't write. And we should
ensure that this "read" section actually completes before this_cpu_dec().

quoted

+	/*
+	 * We never allow heterogeneous nesting of readers. So it is trivial
+	 * to find out the kind of reader we are, and undo the operation
+	 * done by our corresponding percpu_read_lock().
+	 */
+	if (__this_cpu_read(*pcpu_rwlock->reader_refcnt)) {
+		this_cpu_dec(*pcpu_rwlock->reader_refcnt);
+		smp_wmb(); /* Paired with smp_rmb() in sync_reader() */

Given an smp_mb() above, I don't understand the need for this smp_wmb().
Isn't the idea that if the writer sees ->reader_refcnt decremented to
zero, it also needs to see the effects of the corresponding reader's
critical section?

I am equally confused ;)

OTOH, we can probably aboid any barrier if reader_nested_percpu() == T.

quoted

+static void announce_writer_inactive(struct percpu_rwlock *pcpu_rwlock)
+{
+   unsigned int cpu;
+
+   drop_writer_signal(pcpu_rwlock, smp_processor_id());

Why do we drop ourselves twice?  More to the point, why is it important to
drop ourselves first?

And don't we need mb() _before_ we clear ->writer_signal ?

quoted

+static inline void sync_reader(struct percpu_rwlock *pcpu_rwlock,
+			       unsigned int cpu)
+{
+	smp_rmb(); /* Paired with smp_[w]mb() in percpu_read_[un]lock() */

As I understand it, the purpose of this memory barrier is to ensure
that the stores in drop_writer_signal() happen before the reads from
->reader_refcnt in reader_uses_percpu_refcnt(), thus preventing the
race between a new reader attempting to use the fastpath and this writer
acquiring the lock.  Unless I am confused, this must be smp_mb() rather
than smp_rmb().

And note that before sync_reader() we call announce_writer_active() which
already adds mb() before sync_all_readers/sync_reader, so this rmb() looks
unneeded.

But, at the same time, could you confirm that we do not need another mb()
after sync_all_readers() in percpu_write_lock() ? I mean, without mb(),
can't this reader_uses_percpu_refcnt() LOAD leak into the critical section
protected by ->global_rwlock? Then this LOAD can be re-ordered with other
memory operations done by the writer.



Srivatsa, I think that the code would be more understandable if you kill
the helpers like sync_reader/raise_writer_signal. Perhaps even all "write"
helpers, I am not sure. At least, it seems to me that all barriers should
be moved to percpu_write_lock/unlock. But I won't insist of course, up to
you.

And cosmetic nit... How about

	struct xxx {
		unsigned long	reader_refcnt;
		bool		writer_signal;
	}

	struct percpu_rwlock {
		struct xxx __percpu	*xxx;
		rwlock_t		global_rwlock;
	};

?

This saves one alloc_percpu() and ensures that reader_refcnt/writer_signal
are always in the same cache-line.

Oleg.

`h`	back out one level
`j`	next message in thread
`k`	previous message in thread
`l`	drill in
`Esc`	close help / fold thread tree
`?`	toggle this help