← Back to all jobs

Full Stack AI Systems Engineer

Apple · San Diego · Posted 2026-08-20

Apply on the company site →

Job description

Are you a senior engineer who can keep large, AI-augmented systems running reliably at Apple scale? Apple's Stability Engineering team is looking for a seasoned engineer to join our Core team. We build and operate the platforms, services, and infrastructure that turn crash reports from Apple devices into actionable engineering insights. You'll work on systems where LLMs and agents are already part of the production fabric — evolving them, hardening them, and using AI tools to extend what a small team can deliver. Minimum Qualifications: 5+ years of professional software engineering experience building and operating production systems BS in Computer Science or a related field, or equivalent practical experience Fluent use of AI-assisted development tools (coding agents, code review assistants, etc.) to work effectively at scale Demonstrated experience designing and scaling distributed systems (load balancing, active-active topologies, capacity planning, throughput-bound services) Track record of maintaining and evolving production services — observability, operational controls, incident response, and steady iteration on existing systems Strong full-stack instincts; comfortable spanning data infrastructure, backend services, and the user-facing surfaces that consume them Proven ability to operate independently on ambiguous, open-ended problems where the right answer is not obvious Preferred Qualifications: Experience operating LLM- or agent-based features in production environments over time Experience building or maintaining evaluation harnesses, audit trails, or replay infrastructure for AI systems Background in developer tools, observability, crash/stability analysis, or other operating-system-quality domains Familiarity with one or more of: Ruby on Rails, Node.js/TypeScript, Python for production services Experience working in environments with significant deferred scalability work (capacity-constrained, long-lead-time infrastructure)