ML Framework (MetalLM) Engineer
Apple · Cupertino · Posted 2026-08-24
Job description
Apple’s ML Frameworks (MetalLM) team in GPU, Graphics and Machine Learning works on enabling Apple Intelligence through high-performance, distributed inference of GenAI applications (such as LLMs) on Private Cloud Compute. You will get to work on custom-built server hardware that brings the power and security of Apple silicon to the data center. Team also works on GPU acceleration of ML Training frameworks such as PyTorch and JAX using Metal runtime and device backend. We are looking for engineers with systems background who are deeply passionate about building scalable, efficient, and production-grade solutions tailored for high-throughput GPU execution. Minimum Qualifications: 2+ years of programming and problem-solving experience with C/C++/ObjC Experience with GPU kernel development & optimizations using compute programming models such as Metal, CUDA etc. Experience with system level programming and computer architecture Experience with Distributed training or inference techniques Preferred Qualifications: Experience with graph compilers such as Triton, OpenXLA or LLVM/MLIR is a plus Contributions to an AI framework such as PyTorch, JAX or Tensorflow is a plus Good understanding of machine learning fundamentals