Fetching the paper…
Reading the bibliography…
Lecture slide presentations, a sequence of pages that contain text and figures accompanied by speech, are constructed and presented carefully in order to optimally transfer knowledge to students.
Mental representations: A dual coding approach
Allan Paivio · 1990
Earlier work this paper cites.
Cognitive load theory and the format of instruction
Paul Chandler and John Sweller · 1991
Earlier work this paper cites.
Animations need narrations: An experimental test of a dual-coding hypothesis
Richard E Mayer and Richard B Anderson · 1991
Earlier work this paper cites.
Solving the multiple instance problem with axis-parallel rectangles
Thomas G Dietterich, Richard H Lathrop, and Tomás Lozano-Pérez · 1997
Earlier work this paper cites.
The role of interest in learning from scientific text and illustrations: On the distinction between emotional interest and cognitive interest
Shannon F Harp and Richard E Mayer · 1997
Earlier work this paper cites.
Cognitive principles of multimedia learning: The role of modality and contiguity
Roxana Moreno and Richard E Mayer · 1999
Earlier work this paper cites.
Multimedia learning
Richard E Mayer · 2002
Earlier work this paper cites.
Aids to computer-based multimedia learning
Richard E Mayer and Roxana Moreno · 2002
Earlier work this paper cites.
Working memory: looking back and looking forward
Alan Baddeley · 2003
Earlier work this paper cites.
Cognitive load theory: Instructional implications of the interaction between information structures and cognitive architecture
Fred Paas, Alexander Renkl, and John Sweller · 2004
Earlier work this paper cites.
Cognitive effort during note taking
Annie Piolat, Thierry Olive, and Ronald T Kellogg · 2005
Earlier work this paper cites.
Powerpoint’s power in the classroom: Enhancing students’ self-efficacy and attitudes
Joshua E Susskind · 2005
Earlier work this paper cites.
An overview of the tesseract ocr engine
Ray Smith · 2007
Earlier work this paper cites.
Powerpoint, interactive whiteboards, and the visual culture of technology in schools
Gabriel B Reedy · 2008
Earlier work this paper cites.
A realistic dataset for performance evaluation of document layout analysis
Apostolos Antonacopoulos, David Bridson, Christos Papadopoulos, and Stefan Pletschacher · 2009
Cited alongside, same era.
Information retention from powerpoint™ and traditional lectures
April Savoy, Robert W Proctor, and Gavriel Salvendy · 2009
Cited alongside, same era.
Devise: A deep visual-semantic embedding model
Andrea Frome, Greg S Corrado, Jon Shlens, Samy Bengio, Jeff Dean, Marc’Aurelio Ranzato, and Tomas Mikolov · 2013
Cited alongside, same era.
How the design of presentation slides affects audience comprehension: A case for the assertion-evidence approach
Joanna Garner and Michael Alley · 2013
Cited alongside, same era.
An architecture to develop multimodal educative applications with chatbots
David Griol and Zoraida Callejas · 2013
Cited alongside, same era.
Multi-modal language models for lecture video retrieval
Alexander R Fabbri, Irene Li, Prawat Trairatvorakul, Yijiao He, Wei Tai Ting, Robert Tung, Caitlin Westerfield, and Dragomir R Radev · 2018
Later among the works it cites.
Modeling uncertainty with hedged instance embedding
Seong Joon Oh, Kevin Murphy, Jiyan Pan, Joseph Roth, Florian Schroff, and Andrew Gallagher · 2018
Later among the works it cites.
Temporal lecture video fragmentation using word embeddings
Damianos Galanopoulos and Vasileios Mezaris · 2019
Later among the works it cites.
What should i learn first: Introducing lecturebank for nlp education and prerequisite chain learning
Irene Li, Alexander R Fabbri, Robert R Tung, and Dragomir R Radev · 2019
Later among the works it cites.
Polysemous visual-semantic embedding for cross-modal retrieval
Yale Song and Mohammad Soleymani · 2019
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Huizhong Chen, Matthew Cooper, Dhiraj Joshi, and Bernd Girod · 2014
Cited alongside, same era.
Unifying visual-semantic embeddings with multimodal neural language models
Ryan Kiros, Ruslan Salakhutdinov, and Richard S Zemel · 2014
Cited alongside, same era.
Multi-modal and cross-modal for lecture videos retrieval
Nhu Van Nguyen, Mickal Coustaty, and Jean-Marc Ogier · 2014
Cited alongside, same era.
Categorical reparameterization with gumbel-softmax
Eric Jang, Shixiang Gu, and Ben Poole · 2016
Cited alongside, same era.
A comprehensive survey on cross-modal retrieval
Kaiye Wang, Qiyue Yin, Wei Wang, Shu Wu, and Liang Wang · 2016
Cited alongside, same era.
State-of-the-art speech recognition with sequence-to-sequence models
Chung-Cheng Chiu, Tara N Sainath, Yonghui Wu, Rohit Prabhavalkar, Patrick Nguyen, Zhifeng Chen, Anjuli Kannan, Ron J Weiss, Kanishka Rao, Ekaterina Gonina, et al · 2018
Cited alongside, same era.
Bert: Pre-training of deep bidirectional transformers for language understanding
Jacob Devlin, Ming-Wei Chang, Kenton Lee, and Kristina Toutanova · 2018
Cited alongside, same era.
Later among the works it cites.
Huggingface’s transformers: State-of-the-art natural language processing
Thomas Wolf, Lysandre Debut, Victor Sanh, Julien Chaumond, Clement Delangue, Anthony Moi, Pierric Cistac, Tim Rault, Rémi Louf, Morgan Funtowicz, et al · 2019
Later among the works it cites.
Vlengagement: A dataset of scientific video lectures for evaluating population-based engagement
Sahan Bulathwela, Maria Perez-Ortiz, Emine Yilmaz, and John Shawe-Taylor · 2020
Later among the works it cites.
An image is worth 16x16 words: Transformers for image recognition at scale
Alexey Dosovitskiy, Lucas Beyer, Alexander Kolesnikov, Dirk Weissenborn, Xiaohua Zhai, Thomas Unterthiner, Mostafa Dehghani, Matthias Minderer, Georg Heigold, Sylvain Gelly, et al · 2020
Later among the works it cites.
Connecting vision and language with localized narratives
Jordi Pont-Tuset, Jasper Uijlings, Soravit Changpinyo, Radu Soricut, and Vittorio Ferrari · 2020
Later among the works it cites.
Probabilistic embeddings for cross-modal retrieval
Sanghyuk Chun, Seong Joon Oh, Rafael Sampaio De Rezende, Yannis Kalantidis, and Diane Larlus · 2021
Later among the works it cites.
Vilt: Vision-and-language transformer without convolution or region supervision
Wonjae Kim, Bokyung Son, and Ildoo Kim · 2021
Later among the works it cites.
Text-to-image generation grounded by fine-grained user attention
Jing Yu Koh, Jason Baldridge, Honglak Lee, and Yinfei Yang · 2021
Later among the works it cites.
Learning transferable visual models from natural language supervision
Alec Radford, Jong Wook Kim, Chris Hallacy, Aditya Ramesh, Gabriel Goh, Sandhini Agarwal, Girish Sastry, Amanda Askell, Pamela Mishkin, Jack Clark, et al · 2021
Later among the works it cites.