The present application relates to the field of
artificial intelligence technology and computer
document processing technology, in particular to a presentation generation method and
system based on multi-model cooperation and a storage medium, which receives text and identifies types, generates a presentation outline through a large
language model adapted to the field and a prompt word template, performs content refinement
processing on the outline points, analyzes the semantic features and structural features of the outline using a
deep learning model, quantifies the importance of the points and identifies the
logical relationship, determines the page
layout mode, inputs the semantic features of the outline,
layout geometric constraints and visual style consistency constraints into a condition generation model, generates visual elements, and based on the preset constraint conditions and weights, uses a constraint solving
algorithm to automatically
layout the presentation text content and visual elements to generate a presentation page. The present application realizes deep understanding of text,
intelligent planning of layout, accurate matching of vision and automatic optimization of layout, significantly improving the production efficiency and quality of the presentation.