Olivier PERROTIN

Olivier PERROTIN

Biographie

(à compléter)

Publications / Travaux



60 documents

Articles dans une revue

  • Martin Lenglet, Olivier Perrotin, Gérard Bailly. A closer look at internal representations of end-to-end Text-to-Speech models: How is phonetic and acoustic information encoded?. Computer Speech and Language, 2026, 100 (October), pp.101985. ⟨10.1016/j.csl.2026.101985⟩. ⟨hal-05572190⟩
  • Delphine Charuau, Nathalie Henrich Bernardoni, Silvain Gerber, Olivier Perrotin. Hand gesture realisation of contrastive focus in real-time whisper-to-speech synthesis: Investigating the transfer from implicit to explicit control of intonation. Speech Communication, 2026, 177, pp.103344. ⟨10.1016/j.specom.2025.103344⟩. ⟨hal-05448616⟩
  • Ihab Asaad, Maxime Jacquelin, Olivier Perrotin, Laurent Girin, Thomas Hueber. Is Self-Supervised Learning Enough to Fill in the Gap? A Study on Speech Inpainting. Computer Speech and Language, 2026, 99 (July), pp.101922. ⟨10.1016/j.csl.2025.101922⟩. ⟨hal-04728251⟩
  • Olivier Perrotin, Brooke Stephenson, Silvain Gerber, Gérard Bailly, Simon King. Refining the evaluation of speech synthesis: A summary of the Blizzard Challenge 2023. Computer Speech and Language, 2025, 90 (March), pp.101747. ⟨10.1016/j.csl.2024.101747⟩. ⟨hal-04813569⟩
  • Ian Vince Mcloughlin, Olivier Perrotin, Hamid Sharifzadeh, Jacqui Allen, Yan Song. Automated Assessment of Glottal Dysfunction Through Unified Acoustic Voice Analysis. Journal of Voice, 2022, 36 (6), pp.743-754. ⟨10.1016/j.jvoice.2020.08.032⟩. ⟨hal-02987882⟩
  • Olivier Perrotin, Lionel Feugère, Christophe d’Alessandro. Perceptual equivalence of the Liljencrants-Fant and linear-filter glottal flow models. Journal of the Acoustical Society of America, 2021, 150 (2), pp.1273-1285. ⟨10.1121/10.0005879⟩. ⟨hal-03322875⟩
  • Olivier Perrotin, Ian V. Mcloughlin. Glottal Flow Synthesis for Whisper-to-Speech Conversion. IEEE/ACM Transactions on Audio, Speech and Language Processing, 2020, 28, pp.889-900. ⟨10.1109/TASLP.2020.2971417⟩. ⟨hal-02518246⟩
  • Christophe d’Alessandro, Samuel Delalez, Boris Doval, Lionel Feugère, Olivier Perrotin. Les instruments chanteurs. Acoustique et Techniques : trimestriel d’information des professionnels de l’acoustique, 2019, 89, pp.36-43. ⟨hal-02025861⟩
  • Lionel Feugère, Christophe d’Alessandro, Boris Doval, Olivier Perrotin. Cantor Digitalis: chironomic parametric synthesis of singing. EURASIP Journal on Audio, Speech, and Music Processing, 2017, 22, pp.30. ⟨10.1186/s13636-016-0098-5⟩. ⟨hal-01461822⟩
  • Olivier Perrotin, Christophe d’Alessandro. Seeing, listening, drawing: interferences between sensorimotor modalities in the use of a tablet musical interface. ACM Transactions on Applied Perception, 2016, 14 (2), pp.1 – 19. ⟨10.1145/2990501⟩. ⟨hal-01672241⟩
  • Olivier Perrotin, Christophe d’Alessandro. Target Acquisition vs. Expressive Motion: Dynamic Pitch Warping for Intonation Correction. ACM Transactions on Computer-Human Interaction, 2016, 23 (3), pp.17. ⟨10.1145/2897513⟩. ⟨hal-01672238⟩
  • Christophe d’Alessandro, Lionel Feugère, Sylvain Le Beux, Olivier Perrotin, Albert Rilliard. Drawing melodies: Evaluation of chironomic singing synthesis. Journal of the Acoustical Society of America, 2014, 135 (6), pp.3601-3612. ⟨10.1121/1.4875718⟩. ⟨hal-01621771⟩

Communications dans un congrès

  • Ovsev Beliz Ozkan, Michael Jonas, Thomas Hueber, Olivier Perrotin. Prédiction automatique de la fréquence fondamentale à partir des mouvements articulatoires pour la communication silencieuse. JEP 2026 – 36e Journées d’Études sur la Parole, Jun 2026, Montpellier, France. ⟨hal-05691228⟩
  • Gérard Bailly, Olivier Perrotin. The GIPSA-Lab Text-To-Speech System for the Blizzard Challenge 2025. Blizzard Challenge 2025 – 19th Workshop, Aug 2025, Gröningen, Netherlands. ⟨10.21437/Blizzard.2025-5⟩. ⟨hal-05368789v2⟩
  • Gérard Bailly, Elisabeth André, Erica Cooper, Benjamin R. Cowan, Jens Edlund, et al.. Hot topics in speech synthesis evaluation. SSW 2025 – 13th edition of the Speech Synthesis Workshop, Aug 2025, Leeuwarden, Netherlands. pp.1 – 7, ⟨10.21437/ssw.2025-1⟩. ⟨hal-05304059⟩
  • Maxime Jacquelin, Maëva Garnier, Laurent Girin, Rémy Vincent, Olivier Perrotin. LombardTokenizer: Disentanglement and Control of Vocal Effort in a Neural Speech Codec. Interspeech 2025 – 26th Annual Conference of the International Speech Communication Association, Aug 2025, Rotterdam, Netherlands. pp.5778-5782, ⟨10.21437/Interspeech.2025-1639⟩. ⟨hal-05304010⟩
  • Martin Lenglet, Olivier Perrotin, Gérard Bailly. FastLips: an End-to-End Audiovisual Text-to-Speech System with Lip Features Prediction for Virtual Avatars. Interspeech 2024 – 25th Annual Conference of the International Speech Communication Association, Sep 2024, Kos, Greece. pp.3450-3454, ⟨10.21437/Interspeech.2024-462⟩. ⟨hal-04683663⟩
  • Maxime Jacquelin, Maëva Garnier, Laurent Girin, Rémy Vincent, Olivier Perrotin. Exploration de la représentation multidimensionnelle de paramètres acoustiques unidimensionnels de la parole extraits par des modèles profonds non supervisés.. JEP-TALN-RECITAL 2024 – 35èmes Journées d’Études sur la Parole (JEP 2024) 31ème Conférence sur le Traitement Automatique des Langues Naturelles (TALN 2024) 26ème Rencontre des Étudiants Chercheurs en Informatique pour le Traitement Automatique des Langues (RECITAL 2024), Jul 2024, Toulouse, France. pp.82-91. ⟨hal-04623108⟩
  • Delphine Charuau, Nathalie Henrich Bernardoni, Silvain Gerber, Olivier Perrotin. Peut-on marquer un focus contrastif par le geste manuel en suppléance vocale ?. JEP-TALN-RECITAL 2024 – 35èmes Journées d’Études sur la Parole (JEP 2024) 31ème Conférence sur le Traitement Automatique des Langues Naturelles (TALN 2024) 26ème Rencontre des Étudiants Chercheurs en Informatique pour le Traitement Automatique des Langues (RECITAL 2024), Jul 2024, Toulouse, France. pp.142-152. ⟨hal-04623067⟩
  • Gérard Bailly, Romain Legrand, Martin Lenglet, Frédéric Elisei, Maëva Garnier, et al.. Emotags: Computer-Assisted Verbal Labelling of Expressive Audiovisual Utterances for Expressive Multimodal TTS. LREC-COLING 2024 – Joint International Conference on Computational Linguistics, Language Resources and Evaluation (LREC-COLING 2024), Calzolari, Nicoletta; Kan, Min-Yen; Hoste, Veronique; Lenci, Alessandro; Sakti, Sakriani; Xue, Nianwen, May 2024, Turin, Italy. pp.5689-5695. ⟨hal-04709149⟩
  • Delphine Charuau, Nathalie Henrich Bernardoni, Silvain Gerber, Olivier Perrotin. Combining manual control of intonation with whisper articulation in voice substitution: the case of contrastive focus. ISSP 2024 – 13th International Seminar on Speech Production, May 2024, Autrans, France. ⟨hal-04613408⟩
  • Maxime Jacquelin, Maëva Garnier, Laurent Girin, Rémy Vincent, Olivier Perrotin. Exploring the Multidimensional Representation of Unidimensional Speech Acoustic Parameters Extracted by Deep Unsupervised Models. ICASSPW 2024 – IEEE International Conference on Acoustics, Speech and Signal Processing Workshops, Apr 2024, Seoul (Korea), South Korea. pp.858-862, ⟨10.1109/ICASSPW62465.2024.10669904⟩. ⟨hal-04683650⟩
  • Maxime Jacquelin, Maëva Garnier, Laurent Girin, Rémy Vincent, Olivier Perrotin. Exploring the multidimensional representation of unidimensional speech of acoustic parameters extracted by deep unsupervised models. Journée commune AFIA-TLH / AFCP – “Extraction de connaissances interprétables pour l’étude de la communication parlée”, AFIA-TLH; AFCP, Dec 2023, Avignon, France. ⟨hal-04416200⟩
  • Daniel Hernan Molina Villota, Christophe d’Alessandro, Olivier Perrotin. Dynamic pitch warping for expressive vocal retuning. DAFx23 – 26th International Conference on Digital Audio Effects (DAFx23), Sep 2023, Copenhagen, Denmark. pp.118-125. ⟨hal-04256554⟩
  • Olivier Perrotin, Brooke Stephenson, Silvain Gerber, Gérard Bailly. The Blizzard Challenge 2023. Blizzard Challenge 2023 – 18th Workshop, Aug 2023, Grenoble, France. pp.1-27, ⟨10.21437/Blizzard.2023-1⟩. ⟨hal-04269927⟩
  • Martin Lenglet, Olivier Perrotin, Gérard Bailly. The GIPSA-Lab Text-To-Speech System for the Blizzard Challenge 2023. Blizzard Challenge 2023 – 18th Workshop, Aug 2023, Grenoble, France. pp.34-39, ⟨10.21437/Blizzard.2023-3⟩. ⟨hal-04269935⟩
  • Gérard Bailly, Martin Lenglet, Olivier Perrotin, Esther Klabbers. Advocating for text input in multi-speaker text-to-speech systems. SSW 2023 – 12th ISCA Speech Synthesis Workshop (SSW2023), Gérard Bailly; Olivier Perrotin; Thomas Hueber; Damien Lolive; Nicolas Obin, Aug 2023, Grenoble, France. pp.1-7, ⟨10.21437/SSW.2023-1⟩. ⟨hal-04257685⟩
  • Maxime Jacquelin, Maëva Garnier, Laurent Girin, Rémy Vincent, Olivier Perrotin. Exploring the multidimensional representation of individual speech acoustic parameters extracted by deep unsupervised models. SSW 2023 – 12th ISCA Speech Synthesis Workshop (SSW2023), Aug 2023, Grenoble, France. pp.240-241. ⟨hal-04274170⟩
  • Martin Lenglet, Olivier Perrotin, Gérard Bailly. Local Style Tokens: Fine-Grained Prosodic Representations For TTS Expressive Control. SSW 2023 – 12th ISCA Speech Synthesis Workshop (SSW2023), Gérard Bailly; Olivier Perrotin; Thomas Hueber; Damien Lolive; Nicolas Obin, Aug 2023, Grenoble, France. pp.120-126, ⟨10.21437/SSW.2023-19⟩. ⟨hal-04257713⟩
  • Sanjana Sankar, Denis Beautemps, Frédéric Elisei, Olivier Perrotin, Thomas Hueber. Investigating the dynamics of hand and lips in French Cued Speech using attention mechanisms and CTC-based decoding. Interspeech 2023 – 24th Annual Conference of the International Speech Communication Association, ISCA, Aug 2023, Dublin, Ireland. ⟨hal-04126530⟩
  • Maria-Loulou Hajj, Martin Lenglet, Olivier Perrotin, Gérard Bailly. Comparing NLP solutions for the disambiguation of French heterophonic homographs for end-to-end TTS systems. SPECOM 2022 – 24th International Conference on Speech and Computer (SPECOM), Nov 2022, Kitt Gurugram, India. pp.265-278, ⟨10.1007/978-3-031-20980-2_23⟩. ⟨hal-03858736⟩
  • Martin Lenglet, Olivier Perrotin, Gérard Bailly. Speaking Rate Control of end-to-end TTS Models by Direct Manipulation of the Encoder’s Output Embeddings. Interspeech 2022 – 23rd Annual Conference of the International Speech Communication Association, Sep 2022, Incheon, South Korea. pp.11-15, ⟨10.21437/interspeech.2022-759⟩. ⟨hal-03793220v2⟩
  • Luc Ardaillon, Nathalie Henrich Bernardoni, Olivier Perrotin. Voicing decision based on phonemes classification and spectral moments for whisper-to-speech conversion. Interspeech 2022 – 23rd Annual Conference of the International Speech Communication Association, Sep 2022, Incheon, South Korea. pp.2253-2257, ⟨10.21437/interspeech.2022-10675⟩. ⟨hal-03780337⟩
  • Martin Lenglet, Olivier Perrotin, Gérard Bailly. Modélisation de la Parole avec Tacotron2 : Analyse acoustique et phonétique des plongements de caractère. JEP 2022 – 34e Journées d’Études sur la Parole, Jun 2022, Noirmoutier, France. ⟨hal-03727735⟩
  • Kristen M. Murdaugh, Olivier Perrotin, Bruno Gingras, Christian T. Herbst. Correlating Perceptual and Spectral Aspects of Chiaroscuro in Singing — A Pilot Study. PAS7+ – 7th International Physiology and Acoustics of Singing Conference, May 2022, virtual, Brazil. ⟨hal-03736778⟩
  • Daniel Hernan Molina Villota, Christophe d’Alessandro, Olivier Perrotin. Correction dynamique et adaptative de la justesse en voix chantée. CFA 2022 – 16ème Congrès Français d’Acoustique, Société Française d’Acoustique; Laboratoire de Mécanique et d’Acoustique, Apr 2022, Marseille, France. ⟨hal-03848052⟩
  • Daniel Hernan Molina-Villota, Christophe d’Alessandro, Olivier Perrotin. Correction dynamique et adaptative de la justesse en voix chantée. CFA 2022 – 16ème Congrès Français d’Acoustique, Apr 2022, Marseille, France. ⟨hal-03673868⟩
  • Olivier Perrotin, Hussein El Amouri, Gérard Bailly, Thomas Hueber. Evaluating the Extrapolation Capabilities of Neural Vocoders to Extreme Pitch Values. Interspeech 2021 – 22nd Annual Conference of the International Speech Communication Association, Aug 2021, Brno, Czech Republic. pp.11-15, ⟨10.21437/Interspeech.2021-1547⟩. ⟨hal-03338483⟩
  • Martin Lenglet, Olivier Perrotin, Gérard Bailly. Impact of Segmentation and Annotation in French end-to-end Synthesis. SSW 2021 – 11th ISCA Speech Synthesis Workshop, Aug 2021, Budapest, Hungary. pp.13-18, ⟨10.21437/SSW.2021-3⟩. ⟨hal-03362000⟩
  • Kristen M. Murdaugh, Olivier Perrotin, Bruno Gingras, Christian T. Herbst. Correlating Perceptual and Spectral Aspects of Chiaroscuro in Singing — A Pilot Study. PAVA 2021 – Pan-American Vocology Association Symposium, Aug 2021, Salt Lake City, UT, United States. ⟨hal-03736764⟩
  • Olivier Perrotin, Ian Vince Mcloughlin. GFM-Voc: a tool for analysis and modification of the glottis signal. ICVPB 2020 – 12th International Conference on Voice Physiology and Biomechanics, Dec 2020, Grenoble (en ligne), France. ⟨hal-03338865⟩
  • Jacob J Webber, Olivier Perrotin, Simon King. Hider-Finder-Combiner: An Adversarial Architecture for General Speech Signal Modification. Interspeech 2020 – 21st Annual Conference of the International Speech Communication Association, Oct 2020, Shanghai (Virtual Conf), China. ⟨10.21437/interspeech.2020-2558⟩. ⟨hal-02987956⟩
  • Olivier Perrotin, Ian V. Mcloughlin. GFM-Voc: A real-time voice quality modification system. Interspeech 2019 – 20th Annual Conference of the International Speech Communication Association, Sep 2019, Graz, Austria. pp.3685-3686. ⟨hal-02295470⟩
  • Olivier Perrotin, Ian V. Mcloughlin. A Spectral Glottal Flow Model for Source-filter Separation of Speech. ICASSP 2019 – IEEE International Conference on Acoustics, Speech and Signal Processing, May 2019, Brighton, United Kingdom. pp.7160-7164, ⟨10.1109/ICASSP.2019.8682625⟩. ⟨hal-02106450⟩
  • Christophe d’Alessandro, Lionel Feugere, Olivier Perrotin, Samuel Delalez, Boris Doval. Le contrôle des instruments chanteurs. Congrès Français d’Acoustique, CFA 2018, Apr 2018, Le Havre, France. pp.1249-1255. ⟨hal-02008980⟩
  • Olivier Perrotin, Christophe d’Alessandro. Vocal effort modification for singing synthesis. Annual Conference of the International Speech Communication Association (INTERSPEECH 2016), Sep 2016, San Francisco, United States. pp.1235-1239, ⟨10.21437/Interspeech.2016-1096⟩. ⟨hal-01712564⟩
  • Olivier Perrotin, Christophe d’Alessandro. Quel ajustement de hauteur mélodique pour les instruments de musique numériques ?. Journées d’Informatique Musicale (JIM 2015), May 2015, Montréal, Canada. pp.8. ⟨hal-01712565⟩
  • Olivier Perrotin, Christophe d’Alessandro. Visualizing gestures in the control of a digital musical instrument. 14th International Conference on New Interfaces for Musical Expression (NIME 2014), Jun 2014, London, United Kingdom. pp.605-608. ⟨hal-01712570⟩
  • Nicolas d’Alessandro, Christophe d’Alessandro, Lionel Feugere, Maria Astrinaki, Johnty Wang, et al.. Vox Tactum Meets Chorus Digitalis: Seven Years of Singing Surfaces. International Conference on New Interfaces for Musical Expression (NIME 2013), May 2013, Daejeon, South Korea. ⟨hal-01712667⟩
  • Olivier Perrotin, Christophe d’Alessandro. Adaptive mapping for improved pitch accuracy on touch user interfaces. International Conference on New Interfaces for Musical Expression (NIME 2013), May 2013, Daejeon, South Korea. pp.186-189. ⟨hal-01712571⟩
  • Olivier Perrotin, Christophe d’Alessandro. Chironomie et interface pour les instruments de musique virtuels. Journées des Jeunes Chercheurs en Audition, Acoustique musicale et Signal audio (JJCAAS 2012), 2012, Marseille, France. ⟨hal-01712670⟩

Poster de conférence

  • Martin Lenglet, Olivier Perrotin, Gérard Bailly. A Closer Look at Latent Representations of End-to-end TTS Models. Journée commune AFIA-TLH / AFCP – “Extraction de connaissances interprétables pour l’étude de la communication parlée”, Dec 2023, Avignon, France. . ⟨hal-04269953⟩

Proceedings/Recueil des communications

  • Corinne Fredouille, Maëva Garnier, Olivier Perrotin, Marie Tahon. Journée commune AFIA-TLH / AFCP – “Extraction de connaissances interprétables pour l’étude de la communication parlée”. Journée commune AFIA-TLH / AFCP – “Extraction de connaissances interprétables pour l’étude de la communication parlée”, 2023. ⟨hal-04489273⟩

Chapitres d’ouvrage


Autres publications

  • Maxime Jacquelin, Maëva Garnier, Laurent Girin, Rémy Vincent, Olivier Perrotin. The French Lombard Dataset. 2025, ⟨10.5281/zenodo.15533058⟩. ⟨hal-05312350⟩
  • Olivier Perrotin, Christophe d’Alessandro. Modification d’effort vocal pour la synthèse de voix chantée. 2016. ⟨hal-01712660⟩
  • Lionel Feugere, Christophe d’Alessandro, Boris Doval, Olivier Perrotin. Cantor Digitalis: Interactive Voice Factory and Digital Instrument of Sung Vowels/Semi-Vowels. 2016. ⟨hal-01712654⟩

Pré-publications, Documents de travail

  • Shree Harsha Bokkahalli Satish, Harm Lameris, Olivier Perrotin, Gustav Eje Henter, Éva Székely. Speak Your Mind: The Speech Continuation Task as a Probe of Voice-Based Model Bias. 2025. ⟨hal-05368806⟩
  • Olivier Perrotin, Ian V. Mcloughlin. On the Use of a Spectral Glottal Model for the Source-filter Separation of Speech. 2017. ⟨hal-01971707v2⟩

Thèses

  • Olivier Perrotin. Chanter avec les mains : interfaces chironomiques pour les instruments de musique numériques. Interface homme-machine [cs.HC]. Université Paris Sud – Paris XI, 2015. Français. ⟨NNT : 2015PA112207⟩. ⟨tel-01231209⟩

Contact

field is required
field is required
field is required

11 rue des Mathématiques             

38402 Saint-Martin-d’Hères

Lun – Ven : 7:30 – 19:00

Newsletter du GIPSA-lab

La newsletter GIPSA-lab apporte des informations sur les activités et la vie du laboratoire (À venir).

GIPSA-lab - Service SI-Web © 2026 Tous droits réservés.