conference-paper

Graph Attention Neural Network Distributed Model Training

  • 2022 IEEE World AI IoT Congress (AIIoT)
Research footprint

At a glance

Citations
0
References
44
Comments
0
Paper overview

Öz

The scale of neural language models has been increasing significantly over recent years. As a result, the time complexity of training larger language models and resource utilization has been increasing at a higher rate as well. In this research, we propose a distributed implementation of a Graph Attention Neural Network model with 120 million parameters and train it on a cluster of eight GPUs. We demonstrate three times speedup in model training while keeping the stability of accuracy and loss rates during training and testing compared to single GPU instance training.

Record transparency

Publication details

DOI
10.1109/aiiot54504.2022.9817156
OpenAlex
W4285101301
Document type
conference-paper
Language
EN
Source
2022 IEEE World AI IoT Congress (AIIoT)
Last metadata update
Community

Comments

Oturum Açın to join the discussion.

  1. No comments yet. Start the discussion.