R&D Engineering, Staff Engineer - DevOps (f/m)
at Synopsys
Posted 19 hours ago
No clicks
- Compensation
- Not specified
- City
- Country
- France
Currency: Not specified
**Staff Engineer - DevOps (R&D Engineering, f/m) in France** - Lead cross-functional teams to design, implement, and optimize CI/CD pipelines using tools like GitLab, Jenkins, and Git. - Manage Kubernetes clusters at scale, ensuring high availability and reliability. - Mentor junior engineers, fostering a culture of knowledge-sharing and continuous learning. - Requires 7+ years of experience in DevOps roles, familiarity with infrastructure as code (IaC) tools like Terraform, and proficiency in cloud platforms like AWS or GCP. - Experience in AI/ML workflows and infrastructure is a plus. - This role demands a senior-level individual with strong architectural and problem-solving skills, capable of driving strategic decisions in a dynamic, cutting-edge R&D environment.
- Perform complex development activities that may require extensive analysis in areas including cloud deployment and maintenance, as well as distributed system maintenance and scaling
- Use best practices and evangelize through RFCs and mentoring
- Help scale our processes (release, development environments, CI/CD pipelines)
- Root cause investigation, automated release testing, production incident solving
- Design, deploy, and maintain the platform infrastructure using Kubernetes, Pulumi, and Terraform
- Scale GPU and CPU compute resources to support AI model training and inference workloads that grow unpredictably
- Build monitoring, alerting, and observability tooling that catches issues before customers do, using tools like Prometheus, Grafana, or equivalent
- Enable Engineers to deploy models that reduce simulation time from hours to minutes, directly accelerating product innovation for Synopsys customers
- Scale the platform to handle growing customer demand without degrading performance or reliability
- Reduce mean time to recovery during incidents by building better observability and automated remediation into the platform
- Improve developer velocity by streamlining deployment workflows and eliminating friction in the release process
- Establish infrastructure patterns and best practices that the broader engineering team can adopt and scale with
- Prevent outages before they happen by designing resilient, self-healing systems that degrade gracefully under load
- Mentor engineers across the organization through RFCs, documentation, and pairing sessions that raise the bar on platform thinking
- Software development certification
- 3 years’ experience, including managing complex platforms that leverage the Kubernetes technology
- Advanced troubleshooting skills
- Distributed systems design and operation experience (Kubernetes, Pulumi, NATS, Redis)
- Proficiency in scripting with Bash and programming in Python or TypeScript
- Comfort working independently and owning problems end to end, from definition through deployment and monitoring
- You can explain a complex infrastructure tradeoff to a researcher in two sentences without losing the nuance or talking down to them
- When something breaks in production, you stay calm, gather data, and methodically work the problem instead of guessing and restarting things
- You care about the 'why' behind a request, if someone asks for a new service, you ask what problem they are trying to solve before you start provisioning resources
- You are curious enough to test new tools and pragmatic enough to know when the old tool is still the right answer
- You can work across time zones with a distributed team, which means clear written communication and async collaboration are second nature to you
R&D Engineering, Staff Engineer - DevOps (f/m)
at Synopsys
R&D Engineering, Staff Engineer - DevOps (f/m)
at Synopsys
Posted 19 hours ago
No clicks
- Compensation
- Not specified
- City
- Country
- France
Currency: Not specified
**Staff Engineer - DevOps (R&D Engineering, f/m) in France** - Lead cross-functional teams to design, implement, and optimize CI/CD pipelines using tools like GitLab, Jenkins, and Git. - Manage Kubernetes clusters at scale, ensuring high availability and reliability. - Mentor junior engineers, fostering a culture of knowledge-sharing and continuous learning. - Requires 7+ years of experience in DevOps roles, familiarity with infrastructure as code (IaC) tools like Terraform, and proficiency in cloud platforms like AWS or GCP. - Experience in AI/ML workflows and infrastructure is a plus. - This role demands a senior-level individual with strong architectural and problem-solving skills, capable of driving strategic decisions in a dynamic, cutting-edge R&D environment.
- Perform complex development activities that may require extensive analysis in areas including cloud deployment and maintenance, as well as distributed system maintenance and scaling
- Use best practices and evangelize through RFCs and mentoring
- Help scale our processes (release, development environments, CI/CD pipelines)
- Root cause investigation, automated release testing, production incident solving
- Design, deploy, and maintain the platform infrastructure using Kubernetes, Pulumi, and Terraform
- Scale GPU and CPU compute resources to support AI model training and inference workloads that grow unpredictably
- Build monitoring, alerting, and observability tooling that catches issues before customers do, using tools like Prometheus, Grafana, or equivalent
- Enable Engineers to deploy models that reduce simulation time from hours to minutes, directly accelerating product innovation for Synopsys customers
- Scale the platform to handle growing customer demand without degrading performance or reliability
- Reduce mean time to recovery during incidents by building better observability and automated remediation into the platform
- Improve developer velocity by streamlining deployment workflows and eliminating friction in the release process
- Establish infrastructure patterns and best practices that the broader engineering team can adopt and scale with
- Prevent outages before they happen by designing resilient, self-healing systems that degrade gracefully under load
- Mentor engineers across the organization through RFCs, documentation, and pairing sessions that raise the bar on platform thinking
- Software development certification
- 3 years’ experience, including managing complex platforms that leverage the Kubernetes technology
- Advanced troubleshooting skills
- Distributed systems design and operation experience (Kubernetes, Pulumi, NATS, Redis)
- Proficiency in scripting with Bash and programming in Python or TypeScript
- Comfort working independently and owning problems end to end, from definition through deployment and monitoring
- You can explain a complex infrastructure tradeoff to a researcher in two sentences without losing the nuance or talking down to them
- When something breaks in production, you stay calm, gather data, and methodically work the problem instead of guessing and restarting things
- You care about the 'why' behind a request, if someone asks for a new service, you ask what problem they are trying to solve before you start provisioning resources
- You are curious enough to test new tools and pragmatic enough to know when the old tool is still the right answer
- You can work across time zones with a distributed team, which means clear written communication and async collaboration are second nature to you
SIMILAR OPPORTUNITIES
No similar jobs available at the moment.
