Papers / Talks
-
22-052022
PaperImproving Scalability with GPU-Aware Asynchronous Tasks
- Jaemin Choi
- David F. Richards
- Laxmikant Vasudeo Kale
-
21-022021
PosterCharminG: A Scalable GPU-resident Runtime System
- Jaemin Choi
- David F. Richards
- Laxmikant Vasudeo Kale
-
20-042020
PaperAchieving Computation-Communication Overlap with Overdecomposition on GPU Systems
- Jaemin Choi
- David F. Richards
- Laxmikant Vasudeo Kale
-
20-022020
PaperEnd-to-end Performance Modeling of Distributed GPU Applications
- Jaemin Choi
- David F. Richards
- Laxmikant Vasudeo Kale
- Abhinav Bhatele
-
19-062019
PosterACM SRC: Fast Profiling-based Performance Modeling of Distributed GPU Applications
- Jaemin Choi
- Abhinav Bhatele
-
17-112017
PosterACM SRC: Runtime Support for Concurrent Execution of Overdecomposed Heterogeneous Tasks
- Jaemin Choi
- Laxmikant Vasudeo Kale
-
16-142016
PaperRuntime Coordinated Heterogeneous Tasks in Charm++
- Michael P. Robson
- Ronak Akshay Buch
- Laxmikant Vasudeo Kale
-
12-062012
PaperDynamic Scheduling for Work Agglomeration on Heterogeneous Clusters
- Jonathan Lifflander
- G. Carl Evans
- Anshu Arya
- Laxmikant Vasudeo Kale
-
10-162010
PaperScaling Hierarchical N-Body Simulations on GPU Clusters
- Pritish Jetley
- Lukasz Wesolowski
- Filippo Gioachin
- Laxmikant Vasudeo Kale
- Thomas Quinn
-
09-092009
PaperTowards a Framework for Abstracting Accelerators in Parallel Applications: Experience with Cell
- David Kunzman
- Laxmikant Vasudeo Kale
-
09-062009
PaperFlexible Hardware Mapping for Finite Element Simulations on Hybrid CPU / GPU Clusters
- Aaron Becker
- Isaac Dooley
- Laxmikant Vasudeo Kale
-
08-122008
MS Thesis -
06-192006
PosterCharm++ Simplifies Programming for the Cell Processor
- David Kunzman
- Gengbin Zheng
- Eric Bohm
- Laxmikant Vasudeo Kale
-
03-142003
Paper -
98-071998
PaperStatic Networks: A Powerful and Elegant Extension to Concurrent Object-Oriented Languages
- Joshua Yelon
- Laxmikant Vasudeo Kale