Automating Presentation Creation with PASS

Thursday 06 March 2025


The art of presentation-making has long been a tedious task, requiring hours of manual effort and attention to detail. But what if you could automate this process, freeing up your time for more important tasks? A team of researchers has been working on just that, developing a system called PASS (Presentation Automation for Slide Generation and Speech) that can generate high-quality presentation slides and deliver them in an engaging way.


The key innovation behind PASS is its ability to analyze a document and extract the most relevant information, using advanced language models and multimodal processing. This allows it to create slides that are tailored to the specific topic and audience, eliminating the need for manual content creation. But what really sets PASS apart is its capacity to generate speaker notes in real-time, ensuring that the presentation flows smoothly and naturally.


The system’s architecture is based on a modular design, comprising two main modules: Slide Generation and Slide Presentation. The first module takes the input document as its starting point, using an LLM (Large Language Model) to extract relevant topics, content, and points. This information is then used to generate slide titles, bullet points, images, and charts, all of which are carefully designed to be visually appealing and easy to understand.


The second module, Slide Presentation, takes the generated slides and converts them into a spoken narrative using advanced text-to-speech technology. This means that the presenter can simply click through the slides, and the system will generate the corresponding audio script in real-time, allowing for a seamless presentation experience.


But how does PASS perform? The researchers tested their system on a dataset of scientific papers from conferences such as ICML and NeurIPS, comparing its output to existing methods. The results were impressive: PASS outperformed all other systems in terms of coherence, redundancy, and relevance, demonstrating its ability to create high-quality presentations that are both accurate and engaging.


One of the key challenges facing presentation automation is the need to adapt to different audiences and contexts. PASS addresses this issue by using multimodal processing, which allows it to incorporate a wide range of multimedia elements, such as images, videos, and audio clips. This enables the system to create presentations that are tailored to specific audiences and settings, making it an extremely versatile tool.


The potential applications of PASS are vast. Academic researchers could use it to generate presentations for conferences and seminars, freeing up time for more important tasks. Business professionals could use it to create engaging sales pitches and product demos, helping to close deals and drive revenue.


Cite this article: “Automating Presentation Creation with PASS”, The Science Archive, 2025.


Presentation, Automation, Slide Generation, Speech, Language Models, Multimodal Processing, Text-To-Speech, Coherence, Redundancy, Relevance


Reference: Tushar Aggarwal, Aarohi Bhand, “PASS: Presentation Automation for Slide Generation and Speech” (2025).


Leave a Reply