


default search action
19th PPOPP 2014: Orlando, FL, USA
- José E. Moreira, James R. Larus:

ACM SIGPLAN Symposium on Principles and Practice of Parallel Programming, PPoPP '14, Orlando, FL, USA, February 15-19, 2014. ACM 2014, ISBN 978-1-4503-2656-8
Session order 1: opening and conference keynote address
- Mark D. Hill:

21st century computer architecture. 1-2
Session order 2: bugs session
- Tongping Liu, Chen Tian, Ziang Hu, Emery D. Berger

:
PREDATOR: predictive false sharing detection. 3-14 - Paul Thomson, Alastair F. Donaldson

, Adam Betts:
Concurrency testing using schedule bounding: an empirical study. 15-28 - Malavika Samak, Murali Krishna Ramanathan:

Trace driven dynamic deadlock detection and reproduction. 29-42 - Wei-Fan Chiang, Ganesh Gopalakrishnan

, Zvonimir Rakamaric, Alexey Solovyev:
Efficient search for inputs causing high floating-point errors. 43-52
Session order 3: HPC session
- Olivier Tardieu, Benjamin Herta, David Cunningham, David Grove

, Prabhanjan Kambadur, Vijay A. Saraswat, Avraham Shinnar, Mikio Takeuchi, Mandana Vaziri:
X10 and APGAS at Petascale. 53-66 - David Cunningham, David Grove

, Benjamin Herta, Arun Iyengar, Kiyokuni Kawachiya, Hiroki Murata, Vijay A. Saraswat, Mikio Takeuchi, Olivier Tardieu:
Resilient X10: efficient failure-aware programming. 67-80 - Chaoran Yang, Wesley Bland, John M. Mellor-Crummey

, Pavan Balaji:
Portable, MPI-interoperable coarray fortran. 81-92
Session order 4: GPU session
- Yi Yang, Huiyang Zhou

:
CUDA-NP: realizing nested thread-level parallelism in GPGPU applications. 93-106 - Shengen Yan, Chao Li

, Yunquan Zhang, Huiyang Zhou
:
yaSpMV: yet another SpMV framework on GPUs. 107-118 - Michael Bauer, Sean Treichler, Alex Aiken

:
Singe: leveraging warp specialization for high performance on GPUs. 119-130
Session order 5: synchronization session
- Rei Odaira, José G. Castaños, Hisanobu Tomari:

Eliminating global interpreter locks in ruby through hardware transactional memory. 131-142 - Darko Petrovic, Thomas Ropars, André Schiper:

Leveraging hardware message passing for efficient thread synchronization. 143-154 - Maurice Herlihy, Zhiyu Liu:

Well-structured futures and cache locality. 155-166 - Nuno Lourenco Diegues, Paolo Romano

:
Time-warp: lightweight abort minimization in transactional memory. 167-178
Session order 6: PPoPP keynote address
- Kunle Olukotun

:
Beyond parallel programming with domain specific languages. 179-180
Session order 7: algorithms session
- Sukhyun Song, Jeffrey K. Hollingsworth:

Designing and auto-tuning parallel 3-D FFT for computation-communication overlap. 181-192 - Bryan Catanzaro, Alexander Keller, Michael Garland:

A decomposition for in-place matrix transposition. 193-206 - I-Jui Sung, Juan Gómez-Luna

, José María González-Linares
, Nicolás Guil
, Wen-mei W. Hwu:
In-place transposition of rectangular matrices on accelerators. 207-218 - Saeed Maleki, Madanlal Musuvathi, Todd Mytkowicz:

Parallelizing dynamic programming through rank convergence. 219-232
Session order 8: programming systems session
- Sanyam Mehta, Pei-Hung Lin

, Pen-Chung Yew
:
Revisiting loop fusion in the polyhedral framework. 233-246 - Christopher I. Rodrigues, Thomas B. Jablin, Abdul Dakkak, Wen-mei W. Hwu:

Triolet: a programming system that unifies algorithmic skeleton interfaces for high-performance cluster computing. 247-258 - Xu Liu, John M. Mellor-Crummey

:
A tool to analyze the performance of multithreaded programs on NUMA architectures. 259-272
Session order 9: scheduling and determinism session
- Jia Rao, Xiaobo Zhou:

Towards fair and efficient SMP virtual machine scheduling. 273-286 - Kai Lu, Xu Zhou, Tom Bergan, Xiaoping Wang:

Efficient deterministic multithreading without global barriers. 287-300 - Mahdi Eslamimehr, Jens Palsberg:

Race directed scheduling of concurrent programs. 301-314
Session order 10: conference keynote address
- Norm Rubin:

Heterogeneous computing: what does it mean for compiler research? 315-316
Session order 11: non-blocking data structures session
- Aravind Natarajan, Neeraj Mittal:

Fast concurrent lock-free binary search trees. 317-328 - Trevor Brown, Faith Ellen, Eric Ruppert

:
A general technique for non-blocking trees. 329-342 - Dana Drachsler, Martin T. Vechev, Eran Yahav:

Practical concurrent binary search trees via logical ordering. 343-356 - Shahar Timnat, Erez Petrank:

A practical wait-free simulation for lock-free data structures. 357-368
Session order 11: poster session
- Kishore Kumar Pusukuri, Rajiv Gupta, Laxmi Narayan Bhuyan:

Lock contention aware thread migrations. 369-370 - Kyu Hyung Lee

, Dohyeong Kim, Xiangyu Zhang
:
Infrastructure-free logging and replay of concurrent execution on multiple cores. 371-372 - Cfir Aguston, Yosi Ben-Asher, Gadi Haber:

Parallelization hints via code skeletonization. 373-374 - Wenwen Wang, Chenggang Wu, Pen-Chung Yew

, Xiang Yuan, Zhenjiang Wang, Jianjun Li, Xiaobing Feng:
Concurrency bug localization using shared memory access pairs. 375-376 - Vitus J. Leung, David P. Bunde, Jonathan Ebbers, Stefan P. Feer, Nickolas W. Price, Zachary D. Rhodes, Matthew Swank:

Task mapping stencil computations for non-contiguous allocations. 377-378 - Martin Wimmer, Francesco Versaci

, Jesper Larsson Träff, Daniel Cederman, Philippas Tsigas
:
Data structures for task-based priority scheduling. 379-380 - Leonardo Arturo Bautista-Gomez

, Franck Cappello:
Detecting silent data corruption through data dynamic monitoring for scientific applications. 381-382 - Edans F. de O. Sandes

, Guillermo Miranda, Alba Cristina Magalhaes Alves de Melo
, Xavier Martorell
, Eduard Ayguadé
:
Fine-grain parallel megabase sequence comparison with multiple heterogeneous GPUs. 383-384 - Guy Golan-Gueta, G. Ramalingam, Mooly Sagiv, Eran Yahav:

Automatic semantic locking. 385-386 - Ahmed Hassan, Roberto Palmieri

, Binoy Ravindran
:
Optimistic transactional boosting. 387-388 - Kunal Agrawal, Jeremy T. Fineman, Brendan Sheridan, Jim Sukha, Robert Utterback:

Provably good scheduling for parallel programs that use data structures through implicit batching. 389-390 - Lin Ma, Kunal Agrawal, Roger D. Chamberlain:

Theoretical analysis of classic algorithms on highly-threaded many-core GPUs. 391-392 - Daniel Tomkins, Timmie G. Smith, Nancy M. Amato, Lawrence Rauchwerger:

SCCMulti: an improved parallel strongly connected components algorithm. 393-394 - Miao Luo, Xiaoyi Lu, Khaled Hamidouche, Krishna Chaitanya Kandalla, Dhabaleswar K. Panda:

Initial study of multi-endpoint runtime for MPI+OpenMP hybrid programming model on multi-core systems. 395-396 - Katherine E. Isaacs, Todd Gamblin, Abhinav Bhatele, Peer-Timo Bremer

, Martin Schulz
, Bernd Hamann:
Extracting logical structure and identifying stragglers in parallel execution traces. 397-398

manage site settings
To protect your privacy, all features that rely on external API calls from your browser are turned off by default. You need to opt-in for them to become active. All settings here will be stored as cookies with your web browser. For more information see our F.A.Q.


Google
Google Scholar
Semantic Scholar
Internet Archive Scholar
CiteSeerX
ORCID













