Jobs
Indexed directly from employers. Every age is their own publish date.
Indexed directly from employers. Every age is their own publish date.
Indexed directly from employers. Every age is their own publish date.
Searching titles and descriptions for “Architect - GPU Performance”. A match may be a passing mention rather than the job itself. Titles only.
474 roles across 540 listings · show every listing · page 1 of 19
…2 years of experience in LLM training or inference, including performance optimizations, distributed execution, GPU or TPU acceleration, or PyTorch, JAX, or TensorFlow programming…
…2 years of experience in LLM training or inference, including performance optimizations, distributed execution, GPU or TPU acceleration, or PyTorch, JAX, or TensorFlow programming…
…Conduct rigorous ablation studies to optimize model architectures, token budgets, and loss functions. Establish best practices for scaling multimodal training efficiently on large GPU…
…matching frontier performance through training and harness engineering. (blog https://fireworks.ai/blog/open-source-agents-frontier-advisors) - The fine-tuning bottleneck is not…
…Trusted by leading AI labs, startups, and global enterprises, CoreWeave combines superior infrastructure performance with deep technical expertise to accelerate breakthroughs and turn compute…
…Perform end-to-end optimization of AI models, data pipelines, and inference runtimes to enhance performance across current and next-generation GPU architectures. Apply…
…the highest performance in the world for high-performance computing. We are constantly looking for ways to improve our GPU architecture and maintain our…
…Third, you will drive our Performance Dashboards and Observability for cluster-level performance analysis from a stream line telemetry across NICs, Switches, GPUs, and…
…sandboxing technologies, and system-level security architecture. Proven understanding of LLM inference pipelines (Ollama, Llama.cpp, vLLM), GPU-accelerated computing (CUDA, TensorRT), and experience…
…sandboxing technologies, and system-level security architecture. Proven understanding of LLM inference pipelines (Ollama, Llama.cpp, vLLM), GPU-accelerated computing (CUDA, TensorRT), and experience…
…sandboxing technologies, and system-level security architecture. Proven understanding of LLM inference pipelines (Ollama, Llama.cpp, vLLM), GPU-accelerated computing (CUDA, TensorRT), and experience…
…Build and adapt technical demos, sample code, notebooks, benchmark plans, reference architectures, and performance guides. Run deep technical workshops, code labs, architecture reviews, office…
…robotics workloads, including GPU suitability, compute intensity, memory bandwidth, sensor IO, latency, throughput, batching, power, thermal limits, and cost/performance tradeoffs. Ability to engage…
…across GPU/ SOC Architecture teams. A key part of NVIDIA's strength is to innovate in parallel computing fields, delivering the highest performance in…
…Knowledge of GPU architecture, distributed training, high-performance computing, Kubernetes, workload schedulers, cloud infrastructure, or data-center infrastructure. Your base salary will be determined…
…NVIDIA is leading the way in groundbreaking developments in Artificial Intelligence, High-Performance Computing and Visualization. The GPU, our invention, serves as the visual…
…to understand performance bottlenecks and provide quantitative justification for architectural changes or tuning Directly apply computer architecture knowledge of CPUs, GPUs, DSPs, cache coherency…
…to understand performance bottlenecks and provide quantitative justification for architectural changes or tuning Directly apply computer architecture knowledge of CPUs, GPUs, DSPs, cache coherency…
…Your primary responsibility will be to evaluate roughly architected products for thermal feasibility and design thermal solutions for those products. You will also be…
…This role requires equal comfortability discussing GPU cluster architecture and leading multi-million-dollar Cloud commercial negotiations. Key Responsibilities Strategic Customer Ownership Serve as…
…Design and develop designs, architectures, standards, and methods for large-scale distributed systems. Facilitate service capacity planning and demand forecasting, software performance analysis, and…
…Prior exposure to multi-cloud architecture or GPU platforms (AWS Bedrock, Azure AI, GCP Vertex AI, NVIDIA DGX/NGC). Professional level certifications (OCI Architect…
…Lead, inspire, and develop a high-performing engineering organization, ensuring the professional growth of engineers. Architecture & Delivery : Oversee software architecture, development, debugging, and enhancement…
…matching frontier performance through training and harness engineering. (blog https://fireworks.ai/blog/open-source-agents-frontier-advisors) - The fine-tuning bottleneck is not…
…Core Responsibilities Train and fine-tune edge-capable LLMs for tactical defense applications Optimize model performance for various edge hardware (GPU, CPU, and specialized…