Jobs
Indexed directly from employers. Every age is their own publish date.
Indexed directly from employers. Every age is their own publish date.
Indexed directly from employers. Every age is their own publish date.
Searching titles and descriptions for “Senior Engineer – Multimodal ML R&D (Audio, Vision & Language)”. A match may be a passing mention rather than the job itself. Titles only.
15 roles across 17 listings · show every listing
…Industry Keywords Multimodal AI, Generative AI, Vision-Language Models, Audio AI, Edge AI, Vision Transformers Minimum Qualifications: • Bachelor's degree in Engineering, Information Systems…
…Large Language Models, Natural Language Processing, Computer Vision, Speech/Audio Processing, Multimodal AI, Reinforcement Learning, or AI Systems/Infrastructure. Publication track record at top…
…What You'll Do Build the ML and backend systems that turn your vision into a product customers pay for and rely on, going…
…engineers to ensure models are servable from inception - Develop training methodologies for multimodal models that jointly process and generate speech, language, and audio in…
…Strong cross-functional collaboration with hardware teams, software engineering, ML engineering, product, operations, and leadership stakeholders. Excellent judgment in balancing scientific rigor, product urgency…
…Experience working with multimodal foundation models , including vision-language models and models spanning text, image, audio, or video. Experience developing safety, security, or reliability…
…Our platform supports large language, vision, audio, multimodal, and mixture-of-experts models. It gives scientists and engineers the tools to move new optimization…
…persona, and multimodal safety across audio and vision. We are looking for a researcher with a strong track record in applied ML who cares…
…Proven track record of designing, developing, and launching ML models from scratch into production. 2+ years of hands-on experience with vision language models…
…Hands-on experience with ASR, TTS, speech understanding, audio-language models, or multimodal LLMs. Experience building production ML pipelines and MLOps infrastructure. Proven technical…
…Strong cross-functional collaboration with hardware teams, software engineering, ML engineering, product, operations, and leadership stakeholders. Excellent judgment in balancing scientific rigor, product urgency…
…You will evaluate emerging research and industry trends—including advances in large language models, multimodal architectures, and full-duplex natural conversational systems—and translate…
…Open-source contributions in ML or data tooling. Experience with multimodal generation or understanding (vision-language, document AI, video, or audio). Building and optimizing…
…Multimodal GenAI Platform Direction: Lead development and delivery of capabilities spanning large language models, vision-language models, image and video understanding/generation, speech/audio…
…programming skills in Python, Java, Go, or similar languages, with solid software engineering fundamentals ML Fundamentals: a strong grasp of algorithms, from classic statistical…