This position is listed on behalf of a partner company, who manages all applications and next steps. Our partner is looking for an AI Optimization Engineer based in United States. This is a fully remote opportunity focused on improving the performance, scalability, and economics of large-scale AI systems.You will optimize training and inference workloads across the stack, from low-level GPU kernels to distributed infrastructure.The role combines systems engineering, performance analysis, machine learning infrastructure, and compiler-level optimization.Youβll work with modern GPUs and large neural networks, using rigorous measurement and profiling to identify and resolve performance bottlenecks.The position offers the opportunity to influence production AI workloads where improvements in throughput, latency, and cost have meaningful business impact.Youβll collaborate closely with engineering, product, operations, and business teams while contributing to technical direction and engineering standards.As a senior technical contributor, youβll also mentor engineers and help drive a culture of measurable, production-ready optimization.