Jobs
Indexed directly from employers. Every age is their own publish date.
Indexed directly from employers. Every age is their own publish date.
Indexed directly from employers. Every age is their own publish date.
Searching titles and descriptions for “Software Engineer - Voice AI (Inference Runtime)”. A match may be a passing mention rather than the job itself. Titles only.
9 roles
…In-depth knowledge of ML converters/compilers and runtimes, and hardware-accelerated ML inference techniques. Strong understanding of generative AI model architectures and their…
…The work sits at the intersection of distributed systems, AI inference, compilers and runtimes, performance engineering, security, and external partnerships. About the Role We…
…system and runtime that allows engineering teams build and safely, reliably run their AI agents. You will also build own AI Agents that solve…
…the runtime across real, varied hardware, and set patterns and standards for the endpoint codebase. Partner with ML engineers to embed the inference path…
…s voice AI runs on — the partners whose chips, servers, accelerators (CPU/GPU/NPU), cloud and inference platforms, on-device and edge runtimes, and…
…Improve inference performance through quantization, batching, caching, model compilation, runtime tuning, and accelerator-aware optimization. Partner with systems engineers to integrate models into Cloudflare…
COMPANY OVERVIEW Deepgram is the leading platform underpinning the emerging trillion-dollar Voice AI economy, providing real-time APIs for speech-to-text (STT…
…different engineering problem, and it's one of the most important frontiers for bringing voice AI to everyone. As an Embedded AI Engineer, you…
…Strong hands-on experience with AI/ML software, including generative AI models, inference, model serving, or AI application development. Working knowledge of GPU acceleration…