Fetching the paper…
Reading the bibliography…
Augmented Reality (AR) as a platform has the potential to facilitate the reduction of the cocktail party effect.
“DiPCo – Dinner Party Corpus”, 2019
Maarten Segbroeck et al · 1909
Earlier work this paper cites.
“Some Experiments on the Recognition of Speech, with One and with Two Ears”
E. Cherry · 1953
Earlier work this paper cites.
“The generalized correlation method for estimation of time delay”
C. Knapp and G. Carter · 1976
Earlier work this paper cites.
“Voicebox: Speech processing toolbox for matlab”
Mike Brookes · 1997
Earlier work this paper cites.
“Combinatorial Optimization: Algorithms and Complexity”
Christos Papadimitriou and Kenneth Steiglitz · 1998
Earlier work this paper cites.
“The cocktail party phenomenon: A review of research on speech intelligibility in multiple-talker conditions”
Adelbert. Bronkhorst · 2000
Earlier work this paper cites.
Int. Telecommun. Union (ITU), ITU-T Rec. P.862, 2003
“Perceptual evaluation of speech quality (PESQ)” · 2003
Earlier work this paper cites.
“CHiME-6 Challenge: Tackling Multispeaker Speech Recognition for Unsegmented Recordings”, 2020
Shinji Watanabe et al · 2004
Earlier work this paper cites.
“Rescaling Egocentric Vision”, 2020
Dima Damen et al · 2006
Earlier work this paper cites.
“Performance measurement in blind audio source separation”
E. Vincent, R. Gribonval and C. Fevotte · 2006
Earlier work this paper cites.
“A Duality Based Approach for Realtime TV-L1 Optical Flow”
Christopher Zach, Thomas Pock and Horst Bischof · 2007
Earlier work this paper cites.
“COSINE - A corpus of multi-party COnversational Speech In Noisy Environments”
A. Stupakov, E. Hanusa, J. Bilmes and D. Fox · 2009
Cited alongside, same era.
“An algorithm for intelligibility prediction of time–frequency weighted noisy speech”
Cees Taal, Richard Hendriks, Richard Heusdens and Jesper Jensen · 2011
Cited alongside, same era.
“Acoustical capacity as a means of noise control in eating establishments”
Jens Rindel · 2012
Cited alongside, same era.
“The Hearing-Aid Speech Quality Index (HASQI) Version 2”
James. Kates and Kathryn. Arehart · 2014
Cited alongside, same era.
“The Hearing-Aid Speech Perception Index (HASPI)”
James. Kates and Kathryn. Arehart · 2014
Cited alongside, same era.
“ViSQOL: an objective speech quality model”
Andrew Hines, Jan Skoglund, Anil Kokaram and Naomi Harte · 2015
Cited alongside, same era.
“VoxCeleb2: Deep Speaker Recognition”
Joon Chung, Arsha Nagrani and Andrew Zisserman · 2018
Later among the works it cites.
“An Instrumental Intelligibility Metric Based on Information Theory”
Steven Van, W. Kleijn and Richard. Hendriks · 2018
Later among the works it cites.
“An Evaluation of Intrusive Instrumental Intelligibility Metrics”
Steven Van, W. Kleijn and Richard Hendriks · 2018
Later among the works it cites.
“Restaurant acoustics–Verbal communication in eating establishments”
JH Rindel · 2019
Later among the works it cites.
“SDR – Half-baked or Well Done?”
Jonathan Roux, Scott Wisdom, Hakan Erdogan and John. Hershey · 2019
Later among the works it cites.
“EgoCom: A Multi-person Multi-modal Egocentric Communications Dataset”
C. Northcutt, S. Zha, S. Lovegrove and R. Newcombe · 2020
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
“High-speed tracking-by-detection without using image information”
Erik Bochinski, Volker Eiselein and Thomas Sikora · 2017
Cited alongside, same era.
“How far are we from solving the 2d & 3d face alignment problem?(and a dataset of 230,000 3d facial landmarks)”
Adrian Bulat and Georgios Tzimiropoulos · 2017
Cited alongside, same era.
“Scaling Egocentric Vision: The EPIC-KITCHENS Dataset”
Dima Damen et al · 2018
Cited alongside, same era.
“The fifth ‘CHiME’ Speech Separation and Recognition Challenge: Dataset, task and baselines”, 2018
Jon Barker, Shinji Watanabe, Emmanuel Vincent and Jan Trmal · 2018
Cited alongside, same era.
“YOLOv3: An incremental improvement”, 2018
Joseph Redmon and Ali Farhadi · 2018
Cited alongside, same era.
“Retinaface: Single-shot multi-level face localisation in the wild”
Jiankang Deng et al · 2020
Later among the works it cites.
“The Open Images Dataset V4”
Alina Kuznetsova et al · 2020
Later among the works it cites.
“ViSQOL v3: An open source production ready objective speech and audio metric”
Michael Chinen et al · 2020
Later among the works it cites.
“The Hearing-Aid Speech Perception Index (HASPI) Version 2”
James. Kates and Kathryn. Arehart · 2021
Closest in time.
“An Algorithm for Predicting the Intelligibility of Speech Masked by Modulated Noise Maskers”
Jesper Jensen and Cees. Taal · 2022
Closest in time.