Job Description
As a Senior DevOps Engineer, you will:
- Provide extended U.S. coverage for SEV-1 and SEV-2 production incidents during standard business hours.
- Identify manual processes and implement automation solutions that improve efficiency, reduce deployment times, and streamline operational workflows.
- Deploy, manage, and optimize AWS cloud infrastructure using Infrastructure as Code (Terraform) to ensure scalability, security, and reliability.
- Build and maintain monitoring and observability solutions that support application performance, infrastructure health, and production uptime goals.
- Develop, maintain, and optimize CI/CD pipelines using GitHub Actions to enable reliable and efficient software delivery.
- Apply critical security patches and platform updates within established service level objectives while communicating security advisories to stakeholders.
- Create and maintain technical documentation, operational runbooks, and best practices to support knowledge sharing and operational continuity.
Ideal Candidate:
- Nice to Have (WANT) Solid experience working with service mesh (e.g., Istio).
- Security Knowledge: Understanding of DevSecOps principles, including secure deployment practices, vulnerability scanning, and incident response.
- Solid understanding of software development lifecycle (SDLC) and agile delivery (Scrum / Kanban).
- Prior experience in multi-tenant or enterprise-scale platforms.
- Experience with backup/restore automation and disaster recovery procedures.
- Familiarity with Nx monorepo tooling and multi-tenant architectures.
Daily Tasks:
- The primary responsibility of this role is to drive operational efficiency, automation, and reliability of a framework designed to support the web development software lifecycle--from infrastructure provisioning to frontend and backend development, testing, deployment, security compliance, and observability.
- This role focuses on automation, managing infrastructure, monitoring systems, and incident response for the ADAS group.
- The ideal candidate will work closely with Japan-based DevOps team and Japan- and US-based ADAS engineering teams to ensure streamlined development and reliable deployment of ADAS team web applications.
- Weft is an in-house development ecosystem designed to support the entire software lifecycle of web development - from infrastructure provisioning to frontend and backend development, testing, deployment, and observability. It empowers teams to build web applications, microservices, and desktop applications with speed, consistency, and confidence.
Required Skills:
- Required (MUST) 3+ years of professional DevOps / SRE experience.
- CI/CD Pipelines: Proficiency in setting up, maintaining, and troubleshooting CI/CD pipelines (e.g., GitHub Actions).
- Containerization: Solid experience with Docker and Kubernetes, including deployment, scaling, and management.
- Cloud Providers: Hands-on experience with AWS (strongly preferred).
- Strong understanding of IaaS and PaaS offerings, IAM, and networking within cloud environments.
- Infrastructure as Code (IaC): Proficiency with Terraform for managing cloud infrastructure.
- Monitoring & Logging: Experience with monitoring and logging tools (e.g., Prometheus, Grafana, Sentry, ELK stack) for performance tracking and troubleshooting.
- Strong communication skills in cross-functional environments involving engineers, product owners, UX designers, and leadership.