ARTÍCULO
TITULO

Deep Analysis of Job State Statistics on Lomonosov-2 Supercomputer

Dmitry A. Nikitenko    
Vadim V. Voevodin    
Sergey A. Zhumatiy    

Resumen

It is a common knowledge that the increasingly growing capabilities of HPC systems are always limited by a number of efficiency related issues. The reasons can be very different: hardware failures, incorrect job scheduling, peculiarities of algorithm, chosen programming technology specifics, etc. Most of these issues can be detected after precise analysis, but is a very resourceful way to study every application run. Therefore we performed less complicated analysis of the whole supercomputer job flow. In this paper we share our experience of analyzing user applications? job states assigned by the SLURM resource manager that is used on the Lomonosov-2 system at Supercomputing center of Lomonosov Moscow State University. The statistics on job states was collected and it revealed that the ratio of correctly finished jobs (with the COMPLETED state) was rather low. The jobs owners were asked if the distribution of their jobs? states is normal regarding their applications. This user feedback was processed, and some new ways of efficiency gain were revealed as the result.

 Artículos similares

       
 
WoonSeong Jeong, ByungChan Kong and Sang-Guk Yum    
The demand for compact housing is on the rise, driven by the need for floor plans that accommodate stakeholders? preferences. However, clients frequently struggle to convey their spatial needs to professionals, such as architects, due to a lack of means ... ver más
Revista: Applied Sciences

 
Ilia Zaznov, Julian Martin Kunkel, Atta Badii and Alfonso Dufour    
This paper introduces a novel deep learning approach for intraday stock price direction prediction, motivated by the need for more accurate models to enable profitable algorithmic trading. The key problems addressed are effectively modelling complex limi... ver más
Revista: Applied Sciences

 
Chih-Yung Chen, Shang-Feng Lin, Yuan-Wei Tseng, Zhe-Wei Dong and Cheng-Han Cai    
Remote coffee grinder burr wear level assessment system.
Revista: Applied Sciences

 
Tianhao Gao, Meng Zhang, Yifan Zhu, Youjian Zhang, Xiangsheng Pang, Jing Ying and Wenming Liu    
Classifying sports videos is complex due to their dynamic nature. Traditional methods, like optical flow and the Histogram of Oriented Gradient (HOG), are limited by their need for expertise and lack of universality. Deep learning, particularly Convoluti... ver más
Revista: Applied Sciences

 
Xingxing Tong, Ming Chen and Guofu Feng    
The issue of aquatic product quality and safety has gradually become a focal point of societal concern. Analyzing textual comments from people about aquatic products aids in promptly understanding the current sentiment landscape regarding the quality and... ver más
Revista: Applied Sciences