CGO 2011
30 papers
- A HW/SW co-designed heterogeneous multi-core virtual machine for energy-efficient general purpose computing
- A trace-based Java JIT compiler retrofitted from a method-based compiler
- Acculock: Accurate and efficient detection of data races
- Automated locality optimization based on the reuse distance of string operations
- Automated programmable control and parameterization of compiler optimizations
- Automatic parallelization of fine-grained meta-functions on a chip multiprocessor
- Dynamic register promotion of stack variables
- Dynamically accelerating client-side web applications through decoupled execution
- Extendable pattern-oriented optimization directives
- Flow-sensitive pointer analysis for millions of lines of code
- Formally verifying a compiler: Why? How? How far?
- Highly scalable distributed dataflow analysis
- Intel's Array Building Blocks: A retargetable, dynamic compiler and embedded language
- LAR-CC: Large atomic regions with conditional commits
- Language and compiler support for auto-tuning variable-accuracy algorithms
- Link-time optimization for power efficiency in a tagless instruction cache
- MAO - An extensible micro-architectural optimizer
- Neighborhood-aware data locality optimization for NoC-based multicores
- On-chip cache hierarchy-aware tile scheduling for multicore machines
- Phase-based tuning for better utilization of performance-asymmetric multicore processors
- Pinpointing data locality problems using data-centric analysis
- Practical memory checking with Dr. Memory
- Predictive modeling in a polyhedral optimization space
- Prioritizing constraint evaluation for efficient points-to analysis
- Runtime automatic speculative parallelization
- The language, optimizer, and tools mess
- The runtime abort graph and its application to software transactional memory optimization
- Using machines to learn method-specific compilation strategies
- Vapor SIMD: Auto-vectorize once, run everywhere
- Whole-function vectorization