kirancodes.me
To Proof Maintenance & Beyond!
Venues / PPoPP /

PPoPP 2013

45 papers

  1. A peta-scalable CPU-GPU algorithm for global atmospheric simulations · Chao Yang, Wei Xue, Haohuan Fu, Lin Gan, Linfeng Li, Yangtong Xu + 4 more
  2. Adoption protocols for fanout-optimal fault-tolerant termination detection · Jonathan Lifflander, Phil Miller, Laxmikant V. Kalé
  3. Array dataflow analysis for polyhedral X10 programs · Tomofumi Yuki, Paul Feautrier, Sanjay V. Rajopadhye, Vijay A. Saraswat
  4. Automatic problem size sensitive task partitioning on heterogeneous parallel systems · Ivan Grasso, Klaus Kofler, Biagio Cosenza, Thomas Fahringer
  5. Betweenness centrality: algorithms and implementations · Dimitrios Prountzos, Keshav Pingali
  6. Compiler aided manual speculation for high performance concurrent data structures · Lingxiang Xiang, Michael Lee Scott
  7. Complexity analysis and algorithm design for reorganizing data to minimize non-coalesced memory accesses on GPU · Bo Wu, Zhijia Zhao, Eddy Zheng Zhang, Yunlian Jiang, Xipeng Shen
  8. Correct and efficient work-stealing for weak memory models · Nhat Minh Lê, Antoniu Pop, Albert Cohen, Francesco Zappa Nardelli
  9. Data layout optimization for GPGPU architectures · Jun Liu, Wei Ding, Ohyoung Jang, Mahmut T. Kandemir
  10. Data-only flattening for nested data parallelism · Lars Bergstrom, Matthew Fluet, Mike Rainey, John H. Reppy, Stephen Rosen, Adam Shaw
  11. Decomposition techniques for optimal design-space exploration of streaming applications · Shobana Padmanabhan, Yixin Chen, Roger D. Chamberlain
  12. Distributed merge trees · Dmitriy Morozov, Gunther H. Weber
  13. Exploring different automata representations for efficient regular expression matching on GPUs · Xiaodong Yu, Michela Becchi
  14. Expressing graph algorithms using generalized active messages · Nick Edmonds, Jeremiah Willcock, Andrew Lumsdaine
  15. Fast concurrent queues for x86 processors · Adam Morrison, Yehuda Afek
  16. FastLane: improving performance of software transactional memory for low thread counts · Jons-Tobias Wamhoff, Christof Fetzer, Pascal Felber, Etienne Rivière, Gilles Muller
  17. From relational verification to SIMD loop synthesis · Gilles Barthe, Juan Manuel Crespo, Sumit Gulwani, César Kunz, Mark Marron
  18. Ligra: a lightweight graph processing framework for shared memory · Julian Shun, Guy E. Blelloch
  19. Morph algorithms on GPUs · Rupesh Nasre, Martin Burtscher, Keshav Pingali
  20. Multi-level parallel computing of reverse time migration for seismic imaging on blue Gene/Q · Ligang Lu, Karen A. Magerlein
  21. NUMA-aware reader-writer locks · Irina Calciu, David Dice, Yossi Lev, Victor Luchangco, Virendra J. Marathe, Nir Shavit
  22. Online-ABFT: an online algorithm based fault tolerance scheme for soft error detection in iterative methods · Zizhong Chen
  23. Ownership passing: efficient distributed memory programming on multi-core systems · Andrew Friedley, Torsten Hoefler, Greg Bronevetsky, Andrew Lumsdaine, Ching-Chen Ma
  24. Parallel programming with big operators · Changhee Park, Guy L. Steele Jr., Jean-Baptiste Tristan
  25. Parallel schedule synthesis for attribute grammars · Leo A. Meyerovich, Matthew E. Torok, Eric Atkinson, Rastislav Bodík
  26. Parallel suffix array and least common prefix for the GPU · Mrinal Deo, Sean Keely
  27. Programming with hardware lock elision · Yehuda Afek, Amir Levy, Adam Morrison
  28. RaceFree: an efficient multi-threading model for determinism · Kai Lu, Xu Zhou, Xiaoping Wang, Wenzhe Zhang, Gen Li
  29. Reducing contention through priority updates · Julian Shun, Guy E. Blelloch, Jeremy T. Fineman, Phillip B. Gibbons
  30. Relational algorithms for multi-bulk-synchronous processors · Gregory Frederick Diamos, Haicheng Wu, Jin Wang, Ashwin Sanjay Lele, Sudhakar Yalamanchili
  31. Runtime elision of transactional barriers for captured memory · Fernando Miguel Carvalho, João P. Cachopo
  32. Scalable data race detection for partitioned global address space programs · Chang-Seo Park, Koushik Sen, Costin Iancu
  33. Scalable deterministic replay in a parallel full-system emulator · Yufei Chen, Haibo Chen
  34. Scalable statistics counters · Dave Dice, Yossi Lev, Mark Moir
  35. Scheduling parallel programs by work stealing with private deques · Umut A. Acar, Arthur Charguéraud, Mike Rainey
  36. StreamScan: fast scan algorithms for GPUs without global barrier synchronization · Shengen Yan, Guoping Long, Yunquan Zhang
  37. Swift/T: scalable data flow programming for many-task applications · Justin M. Wozniak, Timothy G. Armstrong, Michael Wilde, Daniel S. Katz, Ewing L. Lusk, Ian T. Foster
  38. TeamWork: synchronizing threads globally to detect real deadlocks for multithreaded programs · Yan Cai, Ke Zhai, Shangru Wu, Wing Kwong Chan
  39. The tasks with effects model for safe concurrency · Stephen Heumann, Vikram S. Adve, Shengjie Wang
  40. TigerQuoll: parallel event-based JavaScript · Daniele Bonetta, Walter Binder, Cesare Pautasso
  41. Towards an energy estimator for fault tolerance protocols · Mohammed el Mehdi Diouri, Olivier Glück, Laurent Lefèvre, Franck Cappello
  42. Using hardware transactional memory to correct and simplify and readers-writer lock algorithm · Dave Dice, Yossi Lev, Yujie Liu, Victor Luchangco, Mark Moir
  43. Work-stealing with configurable scheduling strategies · Martin Wimmer, Daniel Cederman, Jesper Larsson Träff, Philippas Tsigas
  44. WuKong: effective diagnosis of bugs at large system scales · Bowen Zhou, Milind Kulkarni, Saurabh Bagchi
  45. ZOOMM: a parallel web browser engine for multicore mobile devices · Calin Cascaval, Seth Fowler, Pablo Montesinos-Ortego, Wayne Piekarski, Mehrdad Reshadi, Behnam Robatmili + 2 more