{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2026,7,16]],"date-time":"2026-07-16T12:00:43Z","timestamp":1784203243589,"version":"3.55.0"},"publisher-location":"New York, NY, USA","reference-count":38,"publisher":"ACM","license":[{"start":{"date-parts":[[2021,2,17]],"date-time":"2021-02-17T00:00:00Z","timestamp":1613520000000},"content-version":"vor","delay-in-days":0,"URL":"https:\/\/www.acm.org\/publications\/policies\/copyright_policy#Background"}],"funder":[{"name":"European Union?s Horizon 2020 Research and Innovation programme","award":["754304"],"award-info":[{"award-number":["754304"]}]}],"content-domain":{"domain":["dl.acm.org"],"crossmark-restriction":true},"short-container-title":[],"published-print":{"date-parts":[[2021,2,17]]},"DOI":"10.1145\/3437801.3441601","type":"proceedings-article","created":{"date-parts":[[2021,2,20]],"date-time":"2021-02-20T23:04:20Z","timestamp":1613862260000},"page":"334-347","update-policy":"https:\/\/doi.org\/10.1145\/crossmark-policy","source":"Crossref","is-referenced-by-count":18,"title":["Advanced synchronization techniques for task-based runtime systems"],"prefix":"10.1145","author":[{"given":"David","family":"\u00c1lvarez","sequence":"first","affiliation":[{"name":"Barcelona Supercomputing Center, Barcelona, Spain"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Kevin","family":"Sala","sequence":"additional","affiliation":[{"name":"Barcelona Supercomputing Center, Barcelona, Spain"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Marcos","family":"Maro\u00f1as","sequence":"additional","affiliation":[{"name":"Barcelona Supercomputing Center, Barcelona, Spain"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Aleix","family":"Roca","sequence":"additional","affiliation":[{"name":"Barcelona Supercomputing Center, Barcelona, Spain"}],"role":[{"vocabulary":"crossref","role":"author"}]},{"given":"Vincen\u00e7","family":"Beltran","sequence":"additional","affiliation":[{"name":"Barcelona Supercomputing Center, Barcelona, Spain"}],"role":[{"vocabulary":"crossref","role":"author"}]}],"member":"320","published-online":{"date-parts":[[2021,2,17]]},"reference":[{"key":"e_1_3_2_1_1_1","doi-asserted-by":"publisher","DOI":"10.26233\/HEALLINK.TUC.30331"},{"key":"e_1_3_2_1_2_1","doi-asserted-by":"publisher","DOI":"10.1145\/378993.379232"},{"key":"e_1_3_2_1_3_1","doi-asserted-by":"publisher","DOI":"10.1145\/1941487.1941507"},{"key":"e_1_3_2_1_4_1","unstructured":"BSC. 2020. Nanos6 GitHub. https:\/\/github.com\/bsc-pm\/nanos6  BSC. 2020. Nanos6 GitHub. https:\/\/github.com\/bsc-pm\/nanos6"},{"key":"e_1_3_2_1_5_1","unstructured":"BSC. 2020. OmpSs-2 Specification. https:\/\/pm.bsc.es\/ftp\/ompss-2\/doc\/spec\/OmpSs-2-Specification.pdf  BSC. 2020. OmpSs-2 Specification. https:\/\/pm.bsc.es\/ftp\/ompss-2\/doc\/spec\/OmpSs-2-Specification.pdf"},{"key":"e_1_3_2_1_6_1","doi-asserted-by":"publisher","DOI":"10.1109\/99.660313"},{"key":"e_1_3_2_1_7_1","unstructured":"Mathieu Desnoyers. 2020. The Common Trace Format. https:\/\/diamon.org\/ctf\/v1.8.3  Mathieu Desnoyers. 2020. The Common Trace Format. https:\/\/diamon.org\/ctf\/v1.8.3"},{"key":"e_1_3_2_1_8_1","doi-asserted-by":"publisher","DOI":"10.1145\/1989493.1989543"},{"key":"e_1_3_2_1_9_1","volume-title":"Euro-Par 2019: Parallel Processing","author":"Dice Dave","unstructured":"Dave Dice and Alex Kogan . 2019. TWA - Ticket Locks Augmented with a Waiting Array . In Euro-Par 2019: Parallel Processing , Ramin Yahyapour (Ed.). Springer International Publishing , Cham , 334--345. Dave Dice and Alex Kogan. 2019. TWA - Ticket Locks Augmented with a Waiting Array. In Euro-Par 2019: Parallel Processing, Ramin Yahyapour (Ed.). Springer International Publishing, Cham, 334--345."},{"key":"e_1_3_2_1_10_1","volume-title":"OpenMP in a New Era of Parallelism, Rudolf Eigenmann and Bronis R","author":"Duran Alejandro","unstructured":"Alejandro Duran , Josep M. Perez , Eduard Ayguad\u00e9 , Rosa M. Badia , and Jesus Labarta . 2008. Extending the OpenMP Tasking Model to Allow Dependent Tasks . In OpenMP in a New Era of Parallelism, Rudolf Eigenmann and Bronis R . de Supinski (Eds.). Springer Berlin Heidelberg , Berlin, Heidelberg , 111--122. Alejandro Duran, Josep M. Perez, Eduard Ayguad\u00e9, Rosa M. Badia, and Jesus Labarta. 2008. Extending the OpenMP Tasking Model to Allow Dependent Tasks. In OpenMP in a New Era of Parallelism, Rudolf Eigenmann and Bronis R. de Supinski (Eds.). Springer Berlin Heidelberg, Berlin, Heidelberg, 111--122."},{"key":"e_1_3_2_1_11_1","volume-title":"2011 38th Annual International Symposium on Computer Architecture (ISCA). 365--376","author":"Esmaeilzadeh H.","unstructured":"H. Esmaeilzadeh , E. Blem , R. S. Amant , K. Sankaralingam , and D. Burger . 2011. Dark silicon and the end of multicore scaling . In 2011 38th Annual International Symposium on Computer Architecture (ISCA). 365--376 . H. Esmaeilzadeh, E. Blem, R. S. Amant, K. Sankaralingam, and D. Burger. 2011. Dark silicon and the end of multicore scaling. In 2011 38th Annual International Symposium on Computer Architecture (ISCA). 365--376."},{"key":"e_1_3_2_1_12_1","unstructured":"Jason Evans. 2020. jemalloc. http:\/\/jemalloc.net\/  Jason Evans. 2020. jemalloc. http:\/\/jemalloc.net\/"},{"key":"e_1_3_2_1_13_1","unstructured":"Ferran Pallar\u00e8s Roca. 2017. Extending OmpSs programming model with task reductions: A compiler and runtime approach. Bachelor's Thesis. Barcelona School of Informatics Universitat Polit\u00e8cnica de Catalunya.  Ferran Pallar\u00e8s Roca. 2017. Extending OmpSs programming model with task reductions: A compiler and runtime approach. Bachelor's Thesis. Barcelona School of Informatics Universitat Polit\u00e8cnica de Catalunya."},{"key":"e_1_3_2_1_14_1","volume-title":"Proceedings of the 2009 linux symposium. Citeseer, 87--93","author":"Fournier Pierre-Marc","year":"2009","unstructured":"Pierre-Marc Fournier , Mathieu Desnoyers , and Michel R Dagenais . 2009 . Combined tracing of the kernel and applications with LTTng . In Proceedings of the 2009 linux symposium. Citeseer, 87--93 . Pierre-Marc Fournier, Mathieu Desnoyers, and Michel R Dagenais. 2009. Combined tracing of the kernel and applications with LTTng. In Proceedings of the 2009 linux symposium. Citeseer, 87--93."},{"key":"e_1_3_2_1_15_1","volume-title":"Evolving OpenMP for Evolving Architectures, Bronis R","author":"Gautier Thierry","unstructured":"Thierry Gautier , Christian Perez , and J\u00e9r\u00f4me Richard . 2018. On the Impact of OpenMP Task Granularity . In Evolving OpenMP for Evolving Architectures, Bronis R . de Supinski, Pedro Valero-Lara, Xavier Martorell, Sergi Mateo Bellido, and Jesus Labarta (Eds.). Springer International Publishing , Cham , 205--221. Thierry Gautier, Christian Perez, and J\u00e9r\u00f4me Richard. 2018. On the Impact of OpenMP Task Granularity. In Evolving OpenMP for Evolving Architectures, Bronis R. de Supinski, Pedro Valero-Lara, Xavier Martorell, Sergi Mateo Bellido, and Jesus Labarta (Eds.). Springer International Publishing, Cham, 205--221."},{"key":"e_1_3_2_1_16_1","volume-title":"OpenMP in the Era of Low Power Devices and Accelerators, Alistair P","author":"Ghosh Priyanka","unstructured":"Priyanka Ghosh , Yonghong Yan , Deepak Eachempati , and Barbara Chapman . 2013. A Prototype Implementation of OpenMP Task Dependency Support . In OpenMP in the Era of Low Power Devices and Accelerators, Alistair P . Rendell, Barbara M. Chapman, and Matthias S. M\u00fcller (Eds.). Springer Berlin Heidelberg , Berlin, Heidelberg , 128--140. Priyanka Ghosh, Yonghong Yan, Deepak Eachempati, and Barbara Chapman. 2013. A Prototype Implementation of OpenMP Task Dependency Support. In OpenMP in the Era of Low Power Devices and Accelerators, Alistair P. Rendell, Barbara M. Chapman, and Matthias S. M\u00fcller (Eds.). Springer Berlin Heidelberg, Berlin, Heidelberg, 128--140."},{"key":"e_1_3_2_1_17_1","unstructured":"GNU Project. 2020. GOMP Source Code. https:\/\/github.com\/gccmirror\/gcc\/tree\/master\/libgomp Accessed: 2020-02-01.  GNU Project. 2020. GOMP Source Code. https:\/\/github.com\/gccmirror\/gcc\/tree\/master\/libgomp Accessed: 2020-02-01."},{"key":"e_1_3_2_1_18_1","doi-asserted-by":"publisher","DOI":"10.1109\/IPDPS.2010.5470425"},{"key":"e_1_3_2_1_19_1","doi-asserted-by":"publisher","DOI":"10.1145\/1810479.1810540"},{"key":"e_1_3_2_1_20_1","doi-asserted-by":"crossref","unstructured":"Ian Karlin Jeff Keasler and Rob Neely. 2013. LULESH 2.0 Updates and Changes. Technical Report LLNL-TR-641973. 1--9 pages.  Ian Karlin Jeff Keasler and Rob Neely. 2013. LULESH 2.0 Updates and Changes. Technical Report LLNL-TR-641973. 1--9 pages.","DOI":"10.2172\/1090032"},{"key":"e_1_3_2_1_21_1","doi-asserted-by":"publisher","DOI":"10.1109\/TPDS.2017.2767046"},{"key":"e_1_3_2_1_22_1","doi-asserted-by":"publisher","DOI":"10.1002\/cpe.1467"},{"key":"e_1_3_2_1_23_1","unstructured":"LLVM Project. 2020. LLVM OpenMP Library Source. https:\/\/github.com\/llvm\/llvm-project\/tree\/master\/openmp  LLVM Project. 2020. LLVM OpenMP Library Source. https:\/\/github.com\/llvm\/llvm-project\/tree\/master\/openmp"},{"key":"e_1_3_2_1_24_1","doi-asserted-by":"publisher","DOI":"10.1109\/HiPC.2019.00053"},{"key":"e_1_3_2_1_25_1","doi-asserted-by":"publisher","DOI":"10.1145\/103727.103729"},{"key":"e_1_3_2_1_26_1","volume-title":"Locality-aware task scheduling and data distribution for OpenMP programs on NUMA systems and manycore processors. Scientific Programming 2015","author":"Muddukrishna Ananya","year":"2015","unstructured":"Ananya Muddukrishna , Peter A Jonsson , and Mats Brorsson . 2015. Locality-aware task scheduling and data distribution for OpenMP programs on NUMA systems and manycore processors. Scientific Programming 2015 ( 2015 ). Ananya Muddukrishna, Peter A Jonsson, and Mats Brorsson. 2015. Locality-aware task scheduling and data distribution for OpenMP programs on NUMA systems and manycore processors. Scientific Programming 2015 (2015)."},{"key":"e_1_3_2_1_27_1","doi-asserted-by":"publisher","DOI":"10.1177\/1094342011434065"},{"key":"e_1_3_2_1_28_1","doi-asserted-by":"publisher","DOI":"10.1109\/IPDPS.2017.69"},{"key":"e_1_3_2_1_29_1","volume-title":"Using and Improving OpenMP for Devices, Tasks, and More, Luiz DeRose, Bronis R","author":"Podobas Artur","unstructured":"Artur Podobas , Mats Brorsson , and Vladimir Vlassov . 2014. TurboB\u0141YSK: Scheduling for Improved Data-Driven Task Performance with Fast Dependency Resolution . In Using and Improving OpenMP for Devices, Tasks, and More, Luiz DeRose, Bronis R . de Supinski, Stephen L. Olivier, Barbara M. Chapman, and Matthias S. M\u00fcller (Eds.). Springer International Publishing , Cham , 45--57. Artur Podobas, Mats Brorsson, and Vladimir Vlassov. 2014. TurboB\u0141YSK: Scheduling for Improved Data-Driven Task Performance with Fast Dependency Resolution. In Using and Improving OpenMP for Devices, Tasks, and More, Luiz DeRose, Bronis R. de Supinski, Stephen L. Olivier, Barbara M. Chapman, and Matthias S. M\u00fcller (Eds.). Springer International Publishing, Cham, 45--57."},{"key":"e_1_3_2_1_30_1","volume-title":"International Workshop on Languages and Compilers for Parallel Computing. Springer, 55--86","author":"Prokopec Aleksandar","year":"2013","unstructured":"Aleksandar Prokopec and Martin Odersky . 2013 . Near optimal work-stealing tree scheduler for highly irregular data-parallel workloads . In International Workshop on Languages and Compilers for Parallel Computing. Springer, 55--86 . Aleksandar Prokopec and Martin Odersky. 2013. Near optimal work-stealing tree scheduler for highly irregular data-parallel workloads. In International Workshop on Languages and Compilers for Parallel Computing. Springer, 55--86."},{"key":"e_1_3_2_1_31_1","doi-asserted-by":"publisher","DOI":"10.1145\/359060.359076"},{"key":"e_1_3_2_1_32_1","doi-asserted-by":"publisher","DOI":"10.1145\/3132747.3132771"},{"key":"e_1_3_2_1_33_1","volume-title":"Towards Data-Flow Parallelization for Adaptive Mesh Refinement Applications. In 2020 IEEE International Conference on Cluster Computing (CLUSTER). IEEE, 314--325","author":"Sala Kevin","year":"2020","unstructured":"Kevin Sala , Alejandro Rico , and Vicen\u00e7 Beltran . 2020 . Towards Data-Flow Parallelization for Adaptive Mesh Refinement Applications. In 2020 IEEE International Conference on Cluster Computing (CLUSTER). IEEE, 314--325 . Kevin Sala, Alejandro Rico, and Vicen\u00e7 Beltran. 2020. Towards Data-Flow Parallelization for Adaptive Mesh Refinement Applications. In 2020 IEEE International Conference on Cluster Computing (CLUSTER). IEEE, 314--325."},{"key":"e_1_3_2_1_35_1","volume-title":"Proceedings of the International Conference for High Performance Computing, Networking, Storage and Analysis","author":"Slaughter Elliott","year":"2020","unstructured":"Elliott Slaughter , Wei Wu , Yuankun Fu , Legends Brandenburg , Nicolai Garcia , Wilhem Kautz , Emily Marx , Kaleb S. Morris , Qinglei Cao , George Bosilca , Seema Mirchandaney , Wonchan Lee , Sean Teichler , Patrick McCormick , and Alex Aiken . 2020 . Task Bench: A Parameterized Benchmark for Evaluating Parallel Runtime Performance . In Proceedings of the International Conference for High Performance Computing, Networking, Storage and Analysis ( Atlanta, Georgia) (SC '20). Association for Computing Machinery. Elliott Slaughter, Wei Wu, Yuankun Fu, Legends Brandenburg, Nicolai Garcia, Wilhem Kautz, Emily Marx, Kaleb S. Morris, Qinglei Cao, George Bosilca, Seema Mirchandaney, Wonchan Lee, Sean Teichler, Patrick McCormick, and Alex Aiken. 2020. Task Bench: A Parameterized Benchmark for Evaluating Parallel Runtime Performance. In Proceedings of the International Conference for High Performance Computing, Networking, Storage and Analysis (Atlanta, Georgia) (SC '20). Association for Computing Machinery."},{"key":"e_1_3_2_1_36_1","doi-asserted-by":"publisher","DOI":"10.1109\/MCSE.2017.29"},{"key":"e_1_3_2_1_37_1","doi-asserted-by":"publisher","DOI":"10.1145\/2541228.2555316"},{"key":"e_1_3_2_1_38_1","doi-asserted-by":"publisher","DOI":"10.1016\/j.procs.2013.05.201"},{"key":"e_1_3_2_1_39_1","volume-title":"An adaptive and hierarchical task scheduling scheme for multi-core clusters. Parallel computing 40, 10","author":"Wang Yizhuo","year":"2014","unstructured":"Yizhuo Wang , Yang Zhang , Yan Su , Xiaojun Wang , Xu Chen , Weixing Ji , and Feng Shi . 2014. An adaptive and hierarchical task scheduling scheme for multi-core clusters. Parallel computing 40, 10 ( 2014 ), 611--627. Yizhuo Wang, Yang Zhang, Yan Su, Xiaojun Wang, Xu Chen, Weixing Ji, and Feng Shi. 2014. An adaptive and hierarchical task scheduling scheme for multi-core clusters. Parallel computing 40, 10 (2014), 611--627."}],"event":{"name":"PPoPP '21: 26th ACM SIGPLAN Symposium on Principles and Practice of Parallel Programming","location":"Virtual Event Republic of Korea","acronym":"PPoPP '21","sponsor":["SIGPLAN ACM Special Interest Group on Programming Languages","SIGHPC ACM Special Interest Group on High Performance Computing, Special Interest Group on High Performance Computing"]},"container-title":["Proceedings of the 26th ACM SIGPLAN Symposium on Principles and Practice of Parallel Programming"],"original-title":[],"link":[{"URL":"https:\/\/dl.acm.org\/doi\/10.1145\/3437801.3441601","content-type":"unspecified","content-version":"vor","intended-application":"text-mining"},{"URL":"https:\/\/dl.acm.org\/doi\/pdf\/10.1145\/3437801.3441601","content-type":"unspecified","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2025,6,17]],"date-time":"2025-06-17T20:17:25Z","timestamp":1750191445000},"score":1,"resource":{"primary":{"URL":"https:\/\/dl.acm.org\/doi\/10.1145\/3437801.3441601"}},"subtitle":[],"short-title":[],"issued":{"date-parts":[[2021,2,17]]},"references-count":38,"alternative-id":["10.1145\/3437801.3441601","10.1145\/3437801"],"URL":"https:\/\/doi.org\/10.1145\/3437801.3441601","relation":{},"subject":[],"published":{"date-parts":[[2021,2,17]]},"assertion":[{"value":"2021-02-17","order":2,"name":"published","label":"Published","group":{"name":"publication_history","label":"Publication History"}}]}}