Systems and techniques that facilitate
semantics-guided domain-specific data augmentation for text-to-graph
parsing are provided. In various embodiments, a
system can access an annotated training dataset, wherein the annotated training dataset can comprise a set of abstract meaning representation graphs respectively corresponding to a set of
natural language sentences. In various aspects, the
system can generate an augmented version of the annotated training dataset, based on applying
semantics-guided composition operations or
semantics-guided substitution operations to the set of abstract meaning representation graphs. In various instances, a
lexicon legend can comprise domain-specific graphs respectively representing discrete tokens unique to
a domain of the annotated training dataset. In some cases, various of the domain-specific graphs can be composed or substituted onto or into various of the set of abstract meaning representation graphs, in response to semantic determinations, such as semantic-type-based determinations, argument-structure-based determinations, or incoming-semantic-relation-based determinations.