- Subject: Re: [PATCH 2/3] Introduce percpu rw semaphores
- From: Eric Dumazet <eric.dumazet@xxxxxxxxx>
- Date: Sat, 28 Jul 2012 22:44:12 +0200
- Cc: Jens Axboe <axboe@xxxxxxxxx>, Andrea Arcangeli <aarcange@xxxxxxxxxx>, Jan Kara <jack@xxxxxxx>, dm-devel@xxxxxxxxxx, linux-kernel@xxxxxxxxxxxxxxx, Jeff Moyer <jmoyer@xxxxxxxxxx>, Alexander Viro <viro@xxxxxxxxxxxxxxxxxx>, kosaki.motohiro@xxxxxxxxxxxxxx, linux-fsdevel@xxxxxxxxxxxxxxx, lwoodman@xxxxxxxxxx, "Alasdair G. Kergon" <agk@xxxxxxxxxx>
- In-reply-to: <Pine.LNX.4.64.1207281240270.30415@file.rdu.redhat.com>
- References: <Pine.LNX.4.64.1206272226050.22857@file.rdu.redhat.com> <20120628111541.GB17515@quack.suse.cz> <Pine.LNX.4.64.1207152051490.4240@file.rdu.redhat.com> <x49ipdmyz4q.fsf@segfault.boston.devel.redhat.com> <Pine.LNX.4.64.1207181512530.10923@file.rdu.redhat.com> <x49k3xzq3jc.fsf@segfault.boston.devel.redhat.com> <Pine.LNX.4.64.1207281236230.30415@file.rdu.redhat.com> <Pine.LNX.4.64.1207281240270.30415@file.rdu.redhat.com>
- Reply-to: device-mapper development <dm-devel@xxxxxxxxxx>
On Sat, 2012-07-28 at 12:41 -0400, Mikulas Patocka wrote:
> Introduce percpu rw semaphores
>
> When many CPUs are locking a rw semaphore for read concurrently, cache
> line bouncing occurs. When a CPU acquires rw semaphore for read, the
> CPU writes to the cache line holding the semaphore. Consequently, the
> cache line is being moved between CPUs and this slows down semaphore
> acquisition.
>
> This patch introduces new percpu rw semaphores. They are functionally
> identical to existing rw semaphores, but locking the percpu rw semaphore
> for read is faster and locking for write is slower.
>
> The percpu rw semaphore is implemented as a percpu array of rw
> semaphores, each semaphore for one CPU. When some thread needs to lock
> the semaphore for read, only semaphore on the current CPU is locked for
> read. When some thread needs to lock the semaphore for write, semaphores
> for all CPUs are locked for write. This avoids cache line bouncing.
>
> Note that the thread that is locking percpu rw semaphore may be
> rescheduled, it doesn't cause bug, but cache line bouncing occurs in
> this case.
>
> Signed-off-by: Mikulas Patocka <mpatocka@xxxxxxxxxx>
I am curious to see how this performs with 4096 cpus ?
Really you shouldnt use rwlock in a path if this might hurt performance.
RCU is probably a better answer.
(bdev->bd_block_size should be read exactly once )
--
dm-devel mailing list
dm-devel@xxxxxxxxxx
https://www.redhat.com/mailman/listinfo/dm-devel
[DM Crypt]
[Fedora Desktop]
[ATA RAID]
[Fedora Marketing]
[Fedora Packaging]
[Fedora SELinux]
[Yosemite Discussion]
[Yosemite Photos]
[KDE Users]
[Fedora Tools]
[Fedora Docs]