Jobs
Indexed directly from employers. Every age is their own publish date.
Indexed directly from employers. Every age is their own publish date.
Indexed directly from employers. Every age is their own publish date.
Searching titles and descriptions for “AI Researcher, On-Device LLM Efficiency”. A match may be a passing mention rather than the job itself. Titles only.
21 roles across 22 listings · show every listing
…You'll surface unmet needs, prototype new tools and features, and collaborate with research, product, and platform to shape the future of AI agent…
…Experience in LLM efficiency research such as efficient attention, inference acceleration, or KV cache compression Experience in on-device AI deployment on mobile or…
…people interact with their devices. On the Siri team we’re solving the challenge of building and shipping powerful LLMs that shape the way…
…services AI & LLM Integration: Manage the evaluation of internally developed and integrated Large Language Models (LLMs), focusing on performance, accuracy, and operational efficiency for…
…Building on years of innovation in intelligent systems and on-device machine learning, we are now scaling efforts in bringing powerful foundation models (on…
…knowledge of on-device algorithm development including hardware-aware ML models and/or optimizing ML compilers for efficient deployment on AI accelerators Proven track…
As Alexa Audio, we own the audio experiences on Alexa enabled devices. These experiences include Music, Podcast, Audio Books, Radio, and Ambient soundscapes. Our…
…reasoning, grounded visual understanding, and language generation within on-device resource budgets Research and implement efficient reasoning techniques — compressed and latent chain-of-thought…
…The LocalAI team is seeking a Systems Software Engineer to build efficient on-device AI software for RTX and DGX-class systems. This role…
…The Opportunity (Summer 2026 AI Internship - Applications Open Now) We're seeking an AI Engineer Intern to work alongside our AI team on large…
…g Iru - Hands-on experience designing and shipping automation with agentic or LLM-based components in an IT or device management context — and the…
…our first party devices and services that combine the best of Google AI, software, and hardware. Teams across this area research, design, and develop…
…in porting models on new Nvidia and AMD GPUs Design, implement, and test functions or components for our AI/DNN/LLM frameworks and tools…
…we build general-purpose AI systems that run efficiently across deployment targets, from data center accelerators to on-device hardware, ensuring low latency, minimal…
…Our AI-powered platform was purpose-built for medical conversations, improving clinical documentation efficiencies while enabling clinicians to focus on what matters most—their…
…AI — all accelerated on NVIDIA GPUs. The Deep Learning Inference team develops and optimizes open-source frameworks that make AI deployment scalable, efficient, and…
…we build general-purpose AI systems that run efficiently across deployment targets, from data center accelerators to on-device hardware, ensuring low latency, minimal…
…A major focus of the role will be the co-development of frontier models and efficient models for Apple silicon and on-device intelligence…
…Publish and open source their cutting-edge AI research. 3. Work on one of the fastest AI supercomputers in the world. 4. Enjoy job…
…Publish and open source their cutting-edge AI research. 3. Work on one of the fastest AI supercomputers in the world. 4. Enjoy job…
…Collaborate with ML and AI researchers to translate cutting-edge models into tangible application features. Implement and optimize the on-device components for LLM…