rwlock different behavior on glibc 2.28 then on previous versions
David Mozes
david.mozes@silk.us
Mon Mar 21 14:01:28 GMT 2022
Hello Christain,
Yes, Thread Sanitizer caught it.
However, we can't run it all the time on the production code.
We need to make a small change of the unlock code to have async programing and backward compatibility for older code.
Thx
David
-----Original Message-----
From: Christian Hoff <christian_hoff@gmx.net>
Sent: Monday, March 21, 2022 8:20 AM
To: David Mozes <david.mozes@silk.us>; Florian Weimer <fweimer@redhat.com>
Cc: libc-help@sourceware.org
Subject: Re: rwlock different behavior on glibc 2.28 then on previous versions
Hallo David,
did you try to run your program with Helgrind already? Helgrind was
designed to catch problems like this. Maybe Thread Sanitizer would also
detect the issue.
Best regards,
Christian
On 3/20/22 5:44 PM, David Mozes wrote:
> Hi,
> Where are we with this?
> Also, the current implementation of the __pthread_rwlock_unlock is problematic.
> On the one hand, it checks the API agreement. However, on the case of violation, the user don't have any induction
> And function return OK
> See example code below:
>
>
> struct threadInfo{
> pthread_t* tid;
> pthread_rwlock_t *rwlock;
> };
>
> void *lockerThreadFunction(void *vargp)
> {
> km_status stat = KM_OK;
> struct threadInfo* Tinfo = (struct threadInfo*)vargp;
>
> do {
> stat = pthread_rwlock_trywrlock(Tinfo->rwlock);
> if (KM_OK == stat) {
> printf("[INFO] - FIRST thread 0x%x locked %p \n", *(Tinfo->tid), Tinfo->rwlock);
> sleep(10);
> stat = pthread_rwlock_unlock(Tinfo->rwlock);
> if (KM_OK == stat) {
> printf("[INFO] FIRST thread 0x%x released %p \n", *(Tinfo->tid), Tinfo->rwlock);
> sleep(1);
> }
> }
> }while(KM_OK != stat);
> return NULL;
> }
>
>
> void *secondThreadFunction(void *vargp)
> {
> km_status stat = KM_OK;
> struct threadInfo* Tinfo = (struct threadInfo*)vargp;
>
> sleep(5);
> do {
> stat = pthread_rwlock_unlock(Tinfo->rwlock);
> if (KM_OK == stat) {
> printf("[INFO] SECOND thread 0x%x released %p \n", *(Tinfo->tid), Tinfo->rwlock);
> sleep(1);
> } else
> printf("[INFO] SECOND thread 0x%x try_to_released %p erro status is: %d \n", *(Tinfo->tid), Tinfo->rwlock,stat);
>
> }while(KM_OK != stat);
> return NULL;
> }
>
> void *thirdThreadFunction(void *vargp)
> {
> km_status stat = KM_OK;
> struct threadInfo* Tinfo = (struct threadInfo*)vargp;
>
> sleep(2);
> do {
> stat = pthread_rwlock_trywrlock(Tinfo->rwlock);
> if (KM_OK == stat) {
> printf("[INFO] THIRD thread 0X%x locked %x\n", *(Tinfo->tid), Tinfo->rwlock);
> sleep(5);
> }
> else {
> printf("[ERROR]: THIRD thread 0x%x failed Locking %p status is %d\n", *(Tinfo->tid), Tinfo->rwlock,stat);
> sleep(1);
> }
> }while(KM_OK != stat);
> return NULL;
> }
>
> void main()
> {
> km_status stat = KM_OK;
> /* km_rwlock lock; */
> pthread_rwlock_t rwlock;
> pthread_t lockerTid = 0, releaseTid = 0, thirdTid=0;
> struct threadInfo firstTinfo, secondTinfo, thirdTinfo;
>
> /* stat = km_rwlock_init(&rwlock, KM_TRUE); */
> stat = pthread_rwlock_init (&rwlock, NULL);
>
> if (KM_OK == stat) {
> printf("read\\writer lock initialized\n");
> printf("Before Thread\n");
>
> firstTinfo.tid = &lockerTid;
> firstTinfo.rwlock = &rwlock;
> secondTinfo.tid = &releaseTid;
> secondTinfo.rwlock = &rwlock;
> thirdTinfo.tid = &thirdTid;
> thirdTinfo.rwlock = &rwlock;
>
>
> pthread_create(&lockerTid, NULL, lockerThreadFunction, &firstTinfo);
> pthread_create(&releaseTid, NULL, secondThreadFunction, &secondTinfo);
> pthread_create(&thirdTid, NULL, thirdThreadFunction, &thirdTinfo);
> printf("After Thread %x\n", lockerTid);
>
> pthread_join(lockerTid, NULL);
> pthread_join(releaseTid, NULL);
> pthread_join(thirdTid, NULL);
>
>
>
> the output we saw is this:
>
>
> the ouput we saw is this:
>
> [root@Rocky85Mozes thread_lock]# ./rw_test
> read\writer lock initialized
> Before Thread
> [INFO] - FIRST thread 0x461f700 locked 0x7ffcda053cc0
> After Thread 461f700
> [ERROR]: THIRD thread 0x361d700 failed Locking 0x7ffcda053cc0 status is 16 -> first thread take the rwlock
> [ERROR]: THIRD thread 0x361d700 failed Locking 0x7ffcda053cc0 status is 16
> [ERROR]: THIRD thread 0x361d700 failed Locking 0x7ffcda053cc0 status is 16
> [INFO] SECOND thread 0x3e1e700 released 0x7ffcda053cc0 -> Second thread unlock the rwlock -- > and get OK!!
> [ERROR]: THIRD thread 0x361d700 failed Locking 0x7ffcda053cc0 status is 16
> [ERROR]: THIRD thread 0x361d700 failed Locking 0x7ffcda053cc0 status is 16
> [ERROR]: THIRD thread 0x361d700 failed Locking 0x7ffcda053cc0 status is 16
> [ERROR]: THIRD thread 0x361d700 failed Locking 0x7ffcda053cc0 status is 16
> [ERROR]: THIRD thread 0x361d700 failed Locking 0x7ffcda053cc0 status is 16
> [INFO] FIRST thread 0x461f700 released 0x7ffcda053cc0
> [ERROR]: THIRD thread 0x361d700 failed Locking 0x7ffcda053cc0 status is 16
>
>
> -----Original Message-----
> From: David Mozes
> Sent: Sunday, March 13, 2022 12:23 PM
> To: Florian Weimer <fweimer@redhat.com>
> Cc: libc-help@sourceware.org
> Subject: RE: rwlock different behavior on glibc 2.28 then on previous versions
>
> Thx Florin,
>
>
> -----Original Message-----
> From: Florian Weimer <fweimer@redhat.com>
> Sent: Friday, March 11, 2022 3:59 PM
> To: David Mozes <david.mozes@silk.us>
> Cc: libc-help@sourceware.org
> Subject: Re: rwlock different behavior on glibc 2.28 then on previous versions
>
> * David Mozes:
>
>> On the previous version of Glibc (before the commit) We could release
>> write lock with different thread as the one that owned the lock (If we
>> like) .
> Sorry, this misuse of the API has always been undefined according to
> POSIX (“Results are undefined if the read-write lock rwlock is not held
> by the calling thread.”).
>
> >> Correct, but practically we can do it on the user's responsibility.
>>> Do you know if this "if" solved some synchronization issue or just enforced the undefined API?
>> This is help very much for async program that has a many threads that
>> communicate with some target. And we like to define a call back
>> function instead of blocking on aio thread that send the data in order
>> to release the rwlock.this improves the performance dramatically .
> Could you use a different synchronization object (perhaps a condition
> variable)?
>
>>> If we use a different synchronization object like semaphores, condition variable (by the way they use mutex under the same API agreement), we will need to write self rwlock , which has less performance than other undefined states.
>>> If we like to keep the API agreement, why not define the different async programming unlock functions?
> Thanks
> Florian
>
> Thx
> David
More information about the Libc-help
mailing list