Computer Scientist II
JOB LEVEL
P40EMPLOYEE ROLE
Individual Contributor- Drive adoption of Infrastructure as Code (Terraform) and enforce platform engineering guidelines and standards
- Define, implement, and continuously improve SLIs, SLOs, and error budgets for Project Graph HTTP APIs and asynchronous compute platforms
- Standardise and optimise Kubernetes-based deployments, scheduling, and resource management for performance and cost efficiency
- Build and enable self-service infrastructure platforms that empower product teams and improve developer productivity
- Lead the adoption of intelligent software solutions (AI agents and MCPs) to reduce operational toil and improve automation at scale
- Define and enforce platform-wide observability standards, covering metrics, logging, tracing, and alerting
- Design, build, and maintain robust observability systems to enable rapid detection, diagnosis, and resolution of issues
- Lead incident response, drive blameless postmortems, and ensure effective follow-through to prevent recurrence
- Improve the reliability, scalability, and performance of asynchronous job scheduling systems built on Kubernetes and Postgres
- Design, maintain, and optimise CI/CD pipelines to ensure fast, safe, and reliable software delivery
- Own data protection and resilience strategies, including backups, restoration testing, and disaster recovery planning
- Architect and implement cloud infrastructure (AWS) aligned to reliability, scalability, performance, and cost optimisation goals
- Reduce operational toil through automation, tooling, and platform improvements, partnering closely with developers to build reliability by design
- Establish and drive guidelines in incident management, reliability engineering, and operational excellence
- Contribute to evaluating system demand, load testing, and performance engineering to ensure systems scale predictably
- Ensure security and compliance guidelines are embedded in platform build and operations (e.g., least privilege, secrets management, secure configurations)
- Participate in an on-call rotation, supporting production systems and driving continuous improvement in operational readiness
- Bachelor’s degree in Computer Science (or equivalent practical experience)
- 8–10 years of experience in Site Reliability Engineering, Platform Engineering, Infrastructure, or Backend Development with a strong operational and production focus
AI‑First & Automation Mindset
- Demonstrated AI-first and automation-first attitude across SRE functions
- Hands-on experience or exposure to AI Agents, MCPs, and AIOps to drive operational efficiency and reduce toil
- Proven ability to identify, prioritise, and eliminate operational toil through automation and intelligent tooling
Cloud, Containers & Platform Engineering
- Deep expertise in Kubernetes in production, including scaling, performance tuning, resolving challenges, and workload optimisation
- Strong experience with containerisation (Docker),Argo and modern deployment patterns
- Hands-on experience designing and operating cloud-native systems on AWS
- Proficiency in Infrastructure as Code (Terraform) and cloud automation
Programming & Systems Engineering
- Strong programming skills in Golang and/or Python, with experience building production-grade systems and tooling
- Experience with Node.js/TypeScript or similar backend technologies is a plus
- Solid understanding of Linux systems, networking fundamentals, REST APIs, and distributed systems design
CI/CD & Developer Productivity
- Strong experience with CI/CD pipelines and tooling (e.g., Argo, Jenkins, Github Actions, CircleCI)
- Ability to design and maintain scalable, reliable build and deployment systems
Observability & Reliability Engineering
- Strong expertise in observability practices, including metrics, logging, tracing, and alerting
- Hands-on experience with tools such as Prometheus, Alertmanager, OpenTelemetry, Jaeger, New Relic, or similar
- Experience defining and operating SLIs, SLOs, and error budgets
- Experience with incident response, production operations, and driving blameless postmortems
Data & Persistence Layer Expertise
- Experience operating production-grade databases (Postgres, MySQL/Aurora, Redis, or similar)
- Strong understanding of data protection, backups, resilience, and disaster recovery (build and testing centered on recovery time objectives and recovery point objectives)
Scalability, Performance & Cost Optimisation
- Experience with forecasting resource needs, performance tuning, and load testing
- Ability to design systems that balance reliability, performance, and cost efficiency
Security & Compliance
- Strong understanding of application and infrastructure security guidelines, including API security, IAM, secrets management, and secure system development
- Familiarity with regulatory and compliance frameworks such as FedRAMP, SOX, and HIPAA, and experience incorporating these requirements into platform design and operations
- Ability to implement and operate systems aligned with security, auditability, and compliance standards in regulated environments
Collaboration & Leadership
- Strong communication skills with the ability to influence engineering teams and drive reliability guidelines
- Proven ability to partner with developers and platform teams to build reliability into systems from the ground up
- Demonstrated ownership approach and ability to drive initiatives end-to-end
Growth Mindset
- Curiosity and ability to learn and adapt to new technologies quickly
- Passion for continuous improvement, operational excellence, and sustainable engineering practices
Internal Opportunities
Creativity, curiosity, and constant learning are celebrated aspects of your career growth journey. We’re glad that you’re pursuing a new opportunity at Adobe!
Put your best foot forward:
1. Update your Resume/CV and Workday profile – don’t forget to include your uniquely ‘Adobe’ experiences and volunteer work.
2. Visit the Internal Mobility page on Inside Adobe to learn more about the process and set up a job alert for roles you’re interested in.
3. Check out these tips to help you prep for interviews.
4. If you are applying for a role outside of your current country, ensure you review the International Resources for Relocating Employees on Inside Adobe, including the impacts to your Benefits, AIP, Equity & Payroll.
Once you apply for a role via Workday, the Talent Team will reach out to you within 2 weeks. If you move into the official interview process with the hiring team, make sure you inform your manager so they can champion your career growth.
At Adobe, you will be immersed in an exceptional work environment that is recognized around the world. You will also be surrounded by colleagues who are committed to helping each other grow through our unique Check-In approach where ongoing feedback flows freely. If you’re looking to make an impact, Adobe's the place for you. Discover what our employees are saying about their career experiences on the Adobe Life blog and explore the meaningful benefits we offer.
Adobe is proud to be an Equal Employment Opportunity employer. We do not discriminate based on gender, race or color, ethnicity or national origin, age, disability, religion, sexual orientation, gender identity or expression, veteran status, or any other applicable characteristics protected by law. Learn more.
Required Skills
Required Languages
🇬🇧 English