Join Google's pioneering Search Platforms team, focusing on LLM Evaluation Infrastructure. This role is instrumental in building the next generation of core infrastructure and benchmarking platforms essential for robust, automated, and scalable LLM evaluations across Google Search.
We are at the forefront of organizing the world's information and making it universally accessible. The LLM Evaluation Infrastructure team is dedicated to this mission, developing advanced systems that power sophisticated AI model assessments. Be part of a team that shapes how billions interact with information and contributes to the evolution of search technology.
As a Senior Software Engineer, you will architect, build, and maintain high-throughput, low-latency platform services and orchestration pipelines for Generative AI (GenAI) Search evaluations. You will design robust systems to monitor metric drift, enhance evaluation throughput, and scale AI-assisted rating workflows efficiently.
Key responsibilities include modernizing and consolidating fragmented evaluation pipelines into a unified platform, developing reusable frameworks, and optimizing system performance for continuous, high-volume evaluation workloads. You will also integrate modern AI tooling and automated validation harnesses to improve developer velocity and system robustness.
We are seeking experienced engineers with a Bachelor's degree or equivalent practical experience. A minimum of 5 years in software development is required, along with 3 years of experience in embedded operating systems, software testing, maintenance, or launching products. Additionally, 1 year of experience in software design and architecture is necessary.
Preferred qualifications include a Master's degree or PhD in Computer Science or a related technical field, 5 years of experience with data structures and algorithms, and 1 year in a technical leadership role. Experience developing accessible technologies is also advantageous.
Technology