PPoPP 2009
43 papers
- A comparison of programming models for multiprocessors with explicitly managed memory hierarchies
- A compiler and runtime system for enabling data mining applications on gpus
- A compiler-directed data prefetching scheme for chip multiprocessors
- A comprehensive strategy for contention management in software transactional memory
- A tunable holistic resiliency approach for high-performance computing systems
- An efficient transactional memory algorithm for computing minimum spanning forest of sparse graphs
- Application-aware management of parallel simulation collections
- Architectural support for cilk computations on many-core architectures
- Atomic quake: using transactional memory in an interactive multiplayer game server
- Backtracking-based load balancing
- Committing conflicting transactions in an STM
- Comparability graph coloring for optimizing utilization of stream register files in stream processors
- Compiler-assisted dynamic scheduling for effective parallelization of loop nests on multicore processors
- Detecting and tolerating asymmetric races
- Effective performance measurement and analysis of multithreaded applications
- Efficient and scalable multiprocessor fair scheduling using distributed weighted round-robin
- Efficient, portable implementation of asynchronous multi-place programs
- Exploiting global optimizations for openmp programs in the openuh compiler
- Formal verification of practical MPI programs
- How much parallelism is there in irregular applications?
- How to build programmable multi-core chips
- Idempotent work stealing
- Industrial perspectives panel
- MPIWiz: subgroup reproducible replay of mpi applications
- Mapping parallelism to multi-cores: a machine learning based approach
- Multi-core demands multi-interfaces
- NePalTM: design and implementation of nested parallelism for transactional memory systems
- OpenMP to GPGPU: a compiler framework for automatic translation and optimization
- Opportunities beyond single-core microprocessors
- Parallel thinking
- Parallelization spectroscopy: analysis of thread-level parallelism in hpc programs
- Petascale computing with accelerators
- Preliminary results on nb-feb, a synchronization primitive for parallel programming
- Safe open-nested transactions through ownership
- Serialization sets: a dynamic dependence-based parallel execution model
- Software transactional distributed shared memory
- Solving dense linear systems on platforms with multiple hardware accelerators
- Stack-based parallel recursion on graphics processors
- Techniques for efficient placement of synchronization primitives
- Topology aware task mapping techniques: an api and case study
- Towards concurrency refactoring for x10
- Transactional memory with strong atomicity using off-the-shelf memory protection hardware
- Turbocharging boosted transactions or: how i learnt to stop worrying and love longer transactions