CGO 2021
35 papers
- ANGHABENCH: A Suite with One Million Compilable C Benchmarks for Code-Size Reduction
- An Experience with Code-Size Optimization for Production iOS Mobile Applications
- An Interval Compiler for Sound Floating-Point Computations
- BuildIt: A Type-Based Multi-stage Programming Framework for Code Generation in C++
- C-for-Metal: High Performance Simd Programming on Intel GPUs
- Cinnamon: A Domain-Specific Language for Binary Profiling and Monitoring
- Compiling Graph Applications for GPU s with GraphIt
- Data Layout and Data Representation Optimizations to Reduce Data Movement Keynote
- ELFies: Executable Region Checkpoints for Performance Analysis and Simulation
- Efficient Execution of Graph Algorithms on CPU with SIMD Extensions
- Enhancing Atomic Instruction Emulation for Cross-ISA Dynamic Binary Translation
- Fine-Grained Pipeline Parallelization for Network Function Programs
- GPA: A GPU Performance Advisor Based on Instruction Sampling
- GoBench: A Benchmark Suite of Real-World Go Concurrency Bugs
- HHVM Jump-Start: Boosting Both Warmup and Steady-State Performance at Scale
- Loop Parallelization using Dynamic Commutativity Analysis
- MLIR: Scaling Compiler Infrastructure for Domain Specific Computation
- Memory-Safe Elimination of Side Channels
- Message from the General Chair
- Message from the Program Chairs
- Object Versioning for Flow-Sensitive Pointer Analysis
- Progressive Raising in Multi-level IR
- Relaxed Peephole Optimization: A Novel Compiler Optimization for Quantum Circuits
- Report from the Artifact Evaluation Committee
- Scaling Up the IFDS Algorithm with Efficient Disk-Assisted Computing
- Seamless Compiler Integration of Variable Precision Floating-Point Arithmetic
- StencilFlow: Mapping Large Stencil Programs to Distributed Spatial Computing Systems
- Thread-Aware Area-Efficient High-Level Synthesis Compiler for Embedded Devices
- Towards a Domain-Extensible Compiler: Optimizing an Image Processing Pipeline on Mobile CPUs
- UNIT: Unifying Tensorized Instruction Compilation
- Unleashing the Low-Precision Computation Potential of Tensor Cores on GPUs
- Variable-Sized Blocks for Locality-Aware SpMV
- Vulkan Vision: Ray Tracing Workload Characterization using Automatic Graphics Instrumentation
- YaskSite: Stencil Optimization Techniques Applied to Explicit ODE Methods on Modern Architectures
- r3d3: Optimized Query Compilation on GPUs