A method, computer program, and computer system for training a graph-to-text generation network is provided. Encoded graph information corresponding to a target sentence is received, and the encoded graph information is decoded based on a biaffine attention score. One or more loss values are determined based on the decoded information, whereby the text-to-graph generation network is trained by minimizing the one or more loss values. A first loss value is generated by reconstructing one or more triple relations based on the biaffine attention score, and a second loss value predicts the graph as a linearized sequence.