git.openfabrics.org - ~shefty/librdmacm.git/commit

author	Sean Hefty <sean.hefty@intel.com>
	Tue, 8 May 2012 00:16:47 +0000 (17:16 -0700)
committer	Sean Hefty <sean.hefty@intel.com>
	Tue, 8 May 2012 00:16:47 +0000 (17:16 -0700)
commit	e20c2f4cc080f49661e1d5f745638105428f52c6
tree	0a763ea65957b893b6af0eba1e31e6cc858869bd	tree \| snapshot
parent	5658ff385e0449a78a325d430163e524b7a97ec4	commit \| diff

rsockets: Optimize synchronization to improve performance

Performance analysis using VTune showed that pthread_mutex_unlock()
is the single biggest contributor to increasing latency for 64-byte
transfers.  Unlocked was followed by get_sw_cqe(), then
__pthread_mutex_lock().  Replace the use of mutexes with an atomic
and a semaphore.  When there's no contention for the lock (which
would usually be the case when using nonblocking sockets), the
code simply increments and decrements an atomic varible.  Semaphores
are only used when contention occurs.

Signed-off-by: Sean Hefty <sean.hefty@intel.com>

src/cma.h		diff \| blob \| history
src/rsocket.c		diff \| blob \| history