Jobs
Indexed directly from employers. Every age is their own publish date.
Indexed directly from employers. Every age is their own publish date.
Indexed directly from employers. Every age is their own publish date.
Searching titles and descriptions for “Machine Learning Engineer - Speech & Multimodal Language Modeling”. A match may be a passing mention rather than the job itself. Titles only.
21 roles across 23 listings · show every listing
…highly skilled Machine Learning Engineer to build and evaluate these experiences, with a specific focus on Multimodal and Speech Language Models. A successful candidate…
…learning methods and machine learning - Experience in building large language models for business application - Experience in building speech recognition, machine translation and natural language…
…Engineering, Machine Learning, or similar technical field. Experience architecting or leading development of full-duplex natural conversational systems, speech-to-speech models, or multimodal…
…learning Experience with Speech LLMs or other Multimodal LLMs Experience with building & deploying AI agents and LLMs Experience with large scale machine learning training…
…You’ll help us advance the state of the art in natural language processing, speech and audio modeling, and multi-modal learning, with a…
…self-supervised learning, synthetic data generation, large language model training, or automatic speech recognition. Experience working with multimodal data (e.g., images, audio, time…
…AI Processing Notice ClickUp may use artificial intelligence and machine learning technologies to help review and screen candidates' employment applications against role-related criteria…
…You will bridge the gap between frontier AI capabilities (large language models, multimodal foundation models, retrieval-augmented generation, speech-to-speech models) and real…
…Modeling and encoding of social signals - Face and body reconstruction and tracking - Large Foundational Models or Multimodal LLMs, such as speech-to-speech LLMs…
…learning-powered AI, enabling breakthroughs in areas like generative AI, computer vision, speech recognition, recommender systems, and large-scale language and multimodal models. Join…
…Multimodal GenAI Platform Direction: Lead development and delivery of capabilities spanning large language models, vision-language models, image and video understanding/generation, speech/audio…
…speech and machine learning techniques - PhD or work experience in a relevant field (CV, Audio, multimodal language models) Experience with distributed training, model compression…
…We do work in all fields of Machine Learning, including, but not limited to, large language models, diffusion models and reinforcement learning, as well…
…We do work in all fields of Machine Learning, including, but not limited to, large language models, diffusion models and reinforcement learning, as well…
…We do work in all fields of Machine Learning, including, but not limited to, large language models, diffusion models and reinforcement learning, as well…
…voice recognition, speech activity detection, Machine learning, Deep learning, Signal processing, speech recognition, Speech enhancement. Minimum Qualifications: • Bachelor's degree in Engineering, Information Systems…
…analysis, machine-learning model development, model validation and serving. - Stay up-to-date with the latest advancements in machine learning, natural language processing, knowledge…
…This covers LLMs (Llama, Phi, Qwen) and multimodal models (vision-language, speech, diffusion), including custom attention, normalization, positional embedding, and modality-specific components. Translate…
…Experience designing machine learning models for sign language. Experience conducting research in Accessibility (e.g., sign language, atypical speech, visual accessibility, or other topics…
…is seeking a Principal Machine Learning Engineer to lead the design, fine-tuning, evaluation, and productionisation of large language models and generative internal AI…
…models Experience with ML areas such as Natural Language Processing, Speech, Multimodal Reasoning & Retrieval, Visual Question & Answering Experience building systems based on machine learning…