kirancodes.me
To Proof Maintenance & Beyond!
Venues / CGO /

CGO 2015

24 papers

  1. A graph-based higher-order intermediate representation · Roland Leißa, Marcel Köster, Sebastian Hack
  2. A parallel abstract interpreter for JavaScript · Kyle Dewey, Vineeth Kashyap, Ben Hardekopf
  3. Approximating flow-sensitive pointer analysis using frequent itemset mining · Vaivaswatha Nagaraj, R. Govindarajan
  4. Automatic data placement into GPU on-chip memory resources · Chao Li, Yi Yang, Zhen Lin, Huiyang Zhou
  5. Branch prediction and the performance of interpreters: don't trust folklore · Erven Rohou, Bharath Narasimha Swamy, André Seznec
  6. Characterizing and enhancing global memory data coalescing on GPUs · Naznin Fauzia, Louis-Noël Pouchet, P. Sadayappan
  7. Checking correctness of code generator architecture specifications · Niranjan Hasabnis, Rui Qiao, R. Sekar
  8. Data provenance tracking for concurrent programs · Brandon Lucia, Luis Ceze
  9. EMEURO: a framework for generating multi-purpose accelerators via deep learning · Lawrence C. McAfee, Kunle Olukotun
  10. Getting in control of your control flow with control-data isolation · William Arthur, Ben Mehne, Reetuparna Das, Todd M. Austin
  11. HELIX-UP: relaxing program semantics to unleash parallelization · Simone Campanoni, Glenn H. Holloway, Gu-Yeon Wei, David M. Brooks
  12. HERMES: a fast cross-ISA binary translator with post-optimization · Xiaochun Zhang, Qi Guo, Yunji Chen, Tianshi Chen, Weiwu Hu
  13. Improving GPGPU energy-efficiency through concurrent kernel execution and DVFS · Qing Jiao, Mian Lu, Huynh Phung Huynh, Tulika Mitra
  14. Locality aware concurrent start for stencil applications · Sunil Shrestha, Guang R. Gao, Joseph B. Manzano, Andrés Márquez, John Feo
  15. Locality-centric thread scheduling for bulk-synchronous programming models on CPU architectures · Hee-Seok Kim, Izzat El Hajj, John A. Stratton, Steven S. Lumetta, Wen-mei W. Hwu
  16. MemorySanitizer: fast detector of uninitialized memory use in C++ · Evgeniy Stepanov, Konstantin Serebryany
  17. On performance debugging of unnecessary lock contentions on multicore processors: a replay-based approach · Long Zheng, Xiaofei Liao, Bingsheng He, Song Wu, Hai Jin
  18. Optimizing and auto-tuning scale-free sparse matrix-vector multiplication on Intel Xeon Phi · Wai Teng Tang, Ruizhe Zhao, Mian Lu, Yun Liang, Huynh Phung Huyng, Xibai Li + 1 more
  19. Optimizing binary translation of dynamically generated code · Byron Hawkins, Brian Demsky, Derek Bruening, Qin Zhao
  20. Optimizing the flash-RAM energy trade-off in deeply embedded systems · James Pallister, Kerstin Eder, Simon J. Hollis
  21. PSLP: padded SLP automatic vectorization · Vasileios Porpodas, Alberto Magni, Timothy M. Jones
  22. Reactive tiling · Jithendra Srinivas, Wei Ding, Mahmut T. Kandemir
  23. Scalable conditional induction variables (CIV) analysis · Cosmin E. Oancea, Lawrence Rauchwerger
  24. Snapshot-based loading-time acceleration for web applications · JinSeok Oh, Soo-Mook Moon