Conference
GeePS: Scalable deep learning on distributed GPUs with a GPU-specialized parameter server
GeePS: Scalable deep learning on distributed GPUs with a GPU-specialized parameter server
On IO Latency Prediction Accuracy and Automated Load Balancing in Consolidated VM Environments
TetriSched: global rescheduling with adaptive plan-ahead in dynamic heterogeneous clusters
Using data transformations for low-latency time series analysis
Springfs: Bridging agility and performance in elastic distributed storage
Toward strong, usable access control for shared distributed data
Specialized storage for big numeric time series
Specialized storage for big numeric time series
Specialized storage for big numeric time series
Visualizing Request-Flow Comparison to Aid Performance Diagnosis in Distributed Systems
Automated diagnosis without predictability is a recipe for failure