Jobs
Indexed directly from employers. Every age is their own publish date.
Indexed directly from employers. Every age is their own publish date.
Indexed directly from employers. Every age is their own publish date.
Searching titles and descriptions for “AI Researcher, On-Device LLM Efficiency”. A match may be a passing mention rather than the job itself. Titles only.
63 roles across 70 listings · show every listing · page 1 of 3
…LLM efficiency research such as efficient attention, inference acceleration, or KV cache compression Experience in on-device AI deployment on mobile or edge devices…
…Amazon Music customers on Alexa/Echo, mobile, and web. Key job responsibilities - Use machine learning, deep learning, LLMs and Agentic AI techniques to create…
…EXAMPLE PROJECTS These are some examples of projects that engineers on our team have worked on recently: - Design and build AI agents for large…
…Our AI-powered platform was purpose-built for medical conversations, improving clinical documentation efficiencies while enabling clinicians to focus on what matters most—their…
…of LLM fine-tuning, alignment, and agentic architectures that operate reliably at scale across many languages and devices, owning delivery from research formulation through…
…execution, or experience in developing and deploying LLMs in production on GPUs, Neuron, TPU or other AI acceleration hardware - Familiarity with recommendation systems is…
Alexa AI is looking for an Applied Scientist to build Alexa+, Amazon's LLM-powered conversational assistant. You will work on key initiatives spanning…
…campaigns based on Customer Retention modeling, effective personalized diagnostic recommendations, LLM Chatbot. You will wield the power of Data and AI to help globally…
…EXAMPLE PROJECTS These are some examples of projects that engineers on our team have worked on recently: - Design and build AI agents for large…
…device backends. You will also partner up with peer science teams to innovate on model quantization and compression techniques for efficient execution on hardware…
…modeling of AI systems or prevailing accelerators/silicon architectures Hands-on proficiency with end-to-end AI hardware architecture or on-device mapping algorithm…
…EXAMPLE PROJECTS These are some examples of projects that engineers on our team have worked on recently: - Design and build AI agents for large…
…LLM-related research projects involving 3-4 researchers and engineers Conduct research and experiments to improve language model accuracy, efficiency, and on-device performance…
…more efficiently and streamline the adoption and embedding of generative AI across Apple. As a Machine Learning Engineer, you will work on building intelligent…
…Optimize agent architecture/orchestration to ensure efficient deployment and operation at scale, with a focus on inference cost optimization. Take ownership of AI quality…
…We are especially looking for PyTorch-focused ML experts driving system-level efficiency from on-device to large-scale models. If you have deep…
…Develop software that interfaces with AI-powered systems and on-device intelligence technologies. Leverage modern developer tools, including AI-assisted coding tools, to improve…
…innovation) - Experience deploying AI/ML models into production systems with direct, verified customer impact - Experience in one or more: NLP/LLMs, graph neural networks…
…class, resource efficient multimodal AI models in support of various perception (vision, audio and speech) based applications for Echo Family of Devices within Amazon…
…and Programming Skills: - Hands-on experience with LLM tools and agents (e.g., Claude, Cline) - Understanding of ML Ops / AI Ops (model lifecycle, deployment…
…MCP utilization and efficiency. AI & Automation Practical experience applying AI and automation to GTM workflows beyond experimentation. Experience working with LLMs (preferably Claude Cowork…
…Address and resolve issues related to AI models optimizations, ensuring high performance and accuracy of AI models. Conduct research on industry trends and innovations…
…AI Research Areas | Intelligence on Devices | Qualcomm Key Responsibilities: Design and develop end-to-end AI tools, models, and software to enable efficient deployment…
…We are a close-knit team of deeply technical AI/ML research scientists who incubate ambitious research ideas and develop new statistical/ML methodologies…
…We also develop state-of-the-art generative AI technologies based on Large Language Models to power innovative features in both Apple’s devices…