ISCA - International Speech
Communication Association


  • Home
  • Post a New Job Offer
<< First  < Prev   1   2   3   4   Next >  Last >> 
  • 2026-08-18 15:01 | Anonymous member

    We are recruiting a fully funded, three-year PhD candidate for a project on Privacy-Preserving Speech Understanding with Multimodal Signals for Clinical Applications, hosted by the MULTISPEECH team at LORIA (Université de Lorraine, Inria and CNRS), in Villers-lès-Nancy, France.

    The project is funded by the AI Grand Est ENACT research chair and will investigate methods to protect speaker identity and sensitive content while preserving the semantic and diagnostic information needed for clinical speech understanding. The work will combine speech and text, with the possibility of incorporating an additional modality such as physiological signals, medical imaging or electronic health-record metadata. It will involve privacy-preserving speech processing, speech/audio foundation models, multimodal machine learning, and evaluation of privacy–utility trade-offs.

    We welcome candidates with a Master’s or engineering degree in computer science, AI, signal or speech processing, applied mathematics, data science, or a related area. Strong Python and machine-learning/deep-learning skills are expected. Experience in speech processing, NLP, privacy-preserving ML, multimodal learning, or AI for healthcare is particularly welcome.

    • Location: Villers-lès-Nancy, France
    • Duration: 3 years, 1 October 2026 to 30 September 2029
    • Start date: 1 October 2026 (or 1 November 2026)
    • Salary: €2,300 gross per month
    • Application deadline: Applications reviewed until the position is filled; final deadline 30 August 2026, potentially extended to 20 September 2026

    Applicants should email a CV, motivation letter, degree transcripts, and, if available, two recommendation letters or the contact details of two referees.

    For the full position description and application details, please see: https://sites.google.com/view/natalia-tomashenko/recruitment-phd-positions Contact: Natalia Tomashenko, natalia.tomashenko@inria.fr

  • 2026-07-31 14:22 | Anonymous member (Administrator)

    Job at Auphonic: Audio Machine/Deep Learning & Signal Processing
    Engineer (Graz/Austria)

    Hallo!

    Auphonic is looking for a Machine/Deep Learning & Signal Processing
    Engineer in Graz, Austria:

    "Work on real-world audio AI — from classic DSP to large deep learning
    models — using our large-scale datasets and GPU infrastructure."

    More details: https://auphonic.com/jobs/ml-engineer

    LG
    Georg

  • 2026-07-29 03:10 | Anonymous member

    The Speech and Multimodal Intelligent Information Processing (SMIIP) Lab at the Chinese University of Hong Kong, Shenzhen has multiple open positions for fully funded Ph.D. student and Postdoc Researchers (2 years contract).

    Our research interests lie in the areas of intelligent speech processing, embodied audition and dialogue system as well as multimodal behavior signal analysis and interpretation.

    Email: mingli369@cuhk.edu.cn

    PI information: Prof. Ming Li https://sai.cuhk.edu.cn/en/teacher/258

    Lab Website: https://smiip-mli.github.io/

    PhD admission information: https://sai.cuhk.edu.cn/en/node/35

    Postdoc recruitment information: https://www.cuhk.edu.cn/en/taxonomy/term/50

    It is highly recommended to contact Prof. Ming Li through email before the application. Thanks.

  • 2026-07-17 10:45 | Anonymous member (Administrator)

    We’re looking for a strong product leader who can drive roadmap, prioritization, and product direction — and who’s excited by the intersection of language, community, and AI. This role will help lead Common Voice, Mozilla’s global, community-powered platform helping make AI more inclusive by ensuring speakers of the world’s languages can be represented in training data.

    We’d especially love to hear from people with:

    • strong senior product leadership experience
    • sound judgement and comfort making tradeoffs across roadmap, platform health, and community needs
    • experience working across engineers, data scientists, designers, linguists, language activists, and other collaborators
    • knowledge of linguistics or languages
    • a basic grounding in computational linguistics or NLP

    Nice to have experience:

    • ML/AI product experience or academic training in ML/AI
    • knowledge of voice or speech AI
    • multilingual or language-centered experience

    Follow this link to apply: https://job-boards.greenhouse.io/mozilla/jobs/7813301 


  • 2026-07-01 06:02 | Anonymous member (Administrator)

    At the MSAD groupe, within the LIST3N laboratory at the University of Technologi of Troyes, we are offering a research internship.

    The aim of this internship is to explore and analyze a new approach based on collaborative knowledge distillation where teacher and student are trained simultaneously.

    Each model mutually enriches the other through a reciprocal influence on their learning processes, going beyond traditional unidirectional transfer.

    To validate this concept, the approach will be applied to speech recognition based on Connectionist Temporal Classification (CTC).

    A known problem with CTC is alignment divergence: models trained separately on the same data often develop inconsistent temporal alignments.

    With collaborative learning, we hypothesize that they will naturally converge towards a unified alignment, thus improving robustness and performance.

    Main Tasks

    • State of the Art : Analyze current research on Collaborative Knowledge Distillation (CKD), with a specific focus on its application to CTC-based speech recognition.
    • Knowledge aggregation: Determine how to efficiently merge knowledge from multiple networks.
    • Learning stabilization: Design regularization mechanisms to ensure a stable and balanced training process.
    • Robustness assessment: Test the performance of the models under various conditions to validate their reliability.

    Candidate's profile

    • Education Level: Master's student (2nd year of research)
    • Skills: Strong background in machine learning and deep learning.
    • Development Requirements : Python proficiency and PyTorch experience.
    • Preferred Assets: Interest in or experience with deep learning, Transformer architecture, speech recognition, and knowledge distillation.

    Intership modalities

    • Internship Start Date : September/October 2026 (depending on student’s availability)
    • Internship Duration : 6 months
    • Location : Université de Technologie de Troyes (UTT), France
    • Language : English or French
    • Expected Internship Level: M1 or M2
    • Gross remuneration : 600€/month

    Application

    If you wish to be considered for this internship opportunity, please send your CV and cover letter to mohammed_faouzi.benzeghiba@utt.fr


  • 2026-04-20 18:10 | Anonymous member (Administrator)

    The University of Oldenburg, Germany, is seeking to fill a permanent
    professorship (salary scale W2) of Medical Physics with focus on
    Technical and Experimental Audiology. For more information about the
    position, please visit https://uol.de/en/job/medical-physics-tea-1027.
    The professorship is part of the Department of Medical Physics and
    Acoustics (https://uol.de/en/mediphysics-acoustics). Please contact
    Prof. Dr. Volker Hohmann (Email: volker.hohmann@uni-oldenburg.de) if you
    have any questions.

  • 2026-04-02 18:40 | Anonymous member

    The Biofeedback Intervention Technology for Speech Lab at NYU (BITS Lab; PI: Tara McAllister), in collaboration with NYU's Music and Audio Research Lab (MARL), is seeking a full-time postdoctoral researcher with expertise in digital signal processing, audio and acoustics, or machine learning and demonstrated experience working with speech data. The project involves developing real-time acoustic biofeedback software for clinical speech intervention, with a focus on automated classification of sibilant productions as accurate or distorted. The postdoc will compare analytic and ML approaches to sibilant classification and contribute to acoustic visualization tools for a web-based clinical application. Candidates from a clinical speech or acoustic phonetics background are also welcome to apply. 

    The position will begin in summer or fall 2026 and is fully remote for U.S.-based candidates (Eastern or Central time zones preferred). In compliance with NYC's Pay Transparency Act, the annual base salary range for this position is $65,000–$70,000. New York University considers factors such as, but not limited to, the specific grant funding and the terms of the research grant when extending an offer. To apply, submit a cover letter, CV, and 2–3 references via Interfolio: https://apply.interfolio.com/183906. Deadline: April 30, 2026.

    BITS Lab is also recruiting PhD students to begin in fall 2027, with an option to be jointly supervised by faculty from MARL. Interested applicants should contact Dr. McAllister directly.

  • 2026-03-27 14:35 | Anonymous member

    We offer a 3-year position (starting, 01.07.2026, fulltime) to a speech scientist or engineer within the Transregional Collaborative Research Center (TRR 318) “Constructing Explainability,” which is jointly run by the Universities of Paderborn and Bielefeld. The TRR investigates how algorithmic transparency can be promoted, particularly in the context of black-box models as used in modern artificial intelligence systems. The position can be used for further academic qualification.

    Our subproject deals with explanatory strategies for recognizing stress in clinical explanatory situations based on multimodal signals (facial expressions and voice). We are developing tools that help clinical staff to better recognize the presence of stress in neurodiverse and neurotypical populations.

    For more details, and a link to the application process, see https://jobs.uni-bielefeld.de/job/view/4874/research-position-m-f-d-in-the-sfb-trr-318-in-the-field-of-phonetics?page_lang=en 

     

  • 2026-03-16 10:26 | Anonymous member

    The Phonetics group at Lancaster University, UK is looking to appoint a Senior Research Associate (i.e. postdoc) in Machine Learning for Speech Processing. The position is available from 1 July 2026 for 18 months.

    The goal is to recover vocal tract movements from the acoustic signal. We are developing ways to integrate physical knowledge into the models, so the inversions are not just accurate but also reveal underlying principles of speech production.

    We're looking for someone with strong ML/speech processing skills who is excited about working at the interface of machine learning, physical modelling, and scientific discovery. The project is funded by The Royal Society and is a collaboration between Sam Kirkham (Lancaster), Anton Ragni (Sheffield) and Aneta Stefanovska (Lancaster).

    More information

  • 2026-02-16 10:10 | Anonymous

    4-year fully PhD funded position, supervised by Prof. Naomi Harte in School of Engineering, Trinity College Dublin, Ireland. Research will explore how multimodal cues used in speech-based interaction can be used to track an active speaker in conversations that go beyond controlled, 2-person scenarios. Full details, including how to apply are here:

    https://www.adaptcentre.ie/careers/phd-studentship-speaker-tracking-in-complex-conversations/

<< First  < Prev   1   2   3   4   Next >  Last >> 
 Organisation  Events   Membership   Help 
 > Board  > Interspeech  > Join - renew  > Sitemap
 > Legal documents  > Workshops  > Membership directory  > Contact
 > Logos      > FAQ
       > Privacy policy

© Copyright 2024 - ISCA International Speech Communication Association - All right reserved.

Powered by Wild Apricot Membership Software