Infrastructure Systems - DevOps Engineer II (Remote in CA only)
Key details
- Work type
- remote
Job Description
Headquarters: 8945 Cal Center Dr, Sacramento, CA 95826, USA URL: http://golden1.com
TITLE: DevOps Engineer IISTATUS: ExemptREPORTS TO: Mgr - DevOps - SREDEPARTMENT: IT - Infrastructure Systems JOB CODE: 12041
PAY RANGE: $123,600.00 - $135,000.00 Annually
POSITION AND PURPOSE
The DevOps Engineer 2 is responsible for leading the automation processes for deploying Infrastructure as Code in both Microsoft Azure and On-Premises environments. The engineer will deploy product updates, identify production issues, and implement integrations that meet our customers’ needs. The ideal candidate will have a solid background in DevOps and Site Reliability Engineering, with significant experience in Terraform, Python, and PowerShell. The engineer will lead the infrastructure-as-code process, manage Linux/Kubernetes cluster environments, and support development teams on API integration strategies. The engineer will design, implement, and optimize CI/CD pipelines for faster and more reliable software releases. Additionally, the engineer will monitor systems, create alerts, and ensure application uptime and performance. Responsibilities also include provisioning and setting up metrics, creating alerts and managing alert suppression, and proposing automation solutions to reduce workload. This role is responsible for implementing and operating cloud platform services and standards defined by Cloud Engineering, with a focus on reliability, security, and scalability.
WHO WE ARE
Golden 1 Credit Union is among the top credit unions in the country. As a member-owned, not-for-profit cooperative, Golden 1 is guided by the credit union philosophy of “people helping people.” We are committed to empowering our members and uplifting our communities as we create a more equitable and financially inclusive California. We welcome all who embrace our Core Values.
WHO YOU ARE
You are an experienced DevOps and Site Reliability Engineering professional who takes ownership of delivering reliable, scalable, and secure platforms in a mission‑critical environment.
You are a hands‑on engineer who designs, implements, and operates automation, Infrastructure as Code, CI/CD pipelines, and observability solutions, and who can independently troubleshoot complex issues and drive them to resolution.
You are a collaborative partner who works closely with development teams, cloud engineering, and IT peers, applies standards consistently, mentors junior engineers, and contributes to a culture of reliability, accountability, and continuous improvement.
THE WORK
GOLDEN 1 RESPONSIBILITES INCLUDE
Independently lead infrastructure-as-code development using Terraform and scripting languages such as Python and PowerShell to support scalable and reliable deployments.
Manage Linux/Kubernetes cluster environments.
Deploy solutions in accordance with Change Management Processes.
Support development teams on API integration strategy and standards development.
Ensure systems are secure against cybersecurity threats.
Identify technical problems and develop software updates and fixes.
Strong Splunk skills for administration, query optimization, alerting, and dashboard development.
Build tools to reduce errors and improve customer experience.
Propose ideas and solutions within the Infrastructure Department to reduce workload through automation.
Design, implement, and optimize CI/CD pipelines for faster and more reliable software releases.
Independently conduct root cause analysis and implement corrective actions.
Design and write tests to investigate infrastructure failure and scaling.
Create and maintain response playbooks across incident management and monitoring tools.
Develop automation to ensure repeatability, eliminate toil, and reduce time to action and repair services.
Analyze key operational metrics to identify opportunities to improve availability.
Implement effective monitoring, alerting, and reduction of alert fatigue.
Manage container orchestration environments and optimize deployment workflows to enhance scalability, reliability, and operational efficiency.
Design, build, and manage containerized environments using Docker.
Create and maintain SLIs, SLOs, and error budgets.
Design and optimize monitoring dashboards and alerting systems to proactively detect and address application performance and uptime issues.
Implement code branching strategies using GitHub functions.
Advanced Terraform syntax and GitLab CI/CD configuration, pipelines, jobs.
Provisioning and setting up metrics in Prometheus, Thanos, and Grafana, creating and managing alerts.
Implement cloud engineering standards, reusable modules, and platform patterns in Microsoft Azure
Operate shared cloud platform services according to Cloud Engineering defined architectures
Ensure infrastructure changes comply with reliability, security, and cost controls established by Cloud Engineering
Maintain operational documentation and runbooks for cloud platform services
QUALIFICATIONS
EDUCATION: Bachelor of science degree (or equivalent) in computer science, engineering, or relevant field.
EXPERIENCE
Over 4 years as a DevOps Engineer in medium to large-scale environments.
Proficient in Windows Server, Linux, and hybrid cloud deployments using Microsoft Azure and VMWare.
Skilled in Git/GitHub workflows, Terraform, Python, PowerShell, and container orchestration (Tanzu, Docker, Kubernetes, OpenShift).
Experienced with CI/CD tools (Jenkins, GitLab CI, Azure DevOps) and observability platforms (Datadog, Prometheus, Grafana, ThousandEyes).
Knowledgeable in log management (ELK Stack) and database technologies (PostgreSQL, MySQL, NoSQL).
Strong background in automating infrastructure provisioning and application deployment using Terraform, Ansible, and Kubernetes.
Proficient in creating and maintaining monitoring dashboards, SLIs, SLOs, and error budgets to ensure application uptime and performance.
Experienced in ensuring infrastructure security, driving automation initiatives, and collaborating across teams to improve reliability and scalability.
Experienced in building observability pipelines and performing advanced queries in log management tools like Splunk for troubleshooting.
Experience implementing and operating Azure-based shared services defined by platform or cloud engineering teams
KNOWLEDGE/SKILLS
Microsoft Azure DevOps Engineer Expert Certification (Required)
Kubernetes Administration Certification (Required)
Linux Certification (Desired)
CORE COMPETENCIES
Takes Initiative – Owns tasks and responsibilities
Delivers Results with Agility – Meets deadlines and adapts
Collaborates Across Teams – Works well within the team solves problems proactively
Handles day-to-day challenges Builds Trust and Credibility – Demonstrates reliability
ORGANIZATIONAL CONTACTS & RELATIONSHIPS
INTERNAL: Regular interaction with Infrastructure Engineers, Computer Operations, IT Programing, Information Security, Network and Storage teams, and IT Service Management staff to support enterprise systems, respond to incidents, and perform scheduled maintenance activities.
EXTERNAL: Interaction with approved technology vendors, hardware and software support providers, and service partners, typically in coordination with IT Systems Manager, for troubleshooting, maintenance, and support escalation.
WORKING CONDITIONS
Work time includes weekend and after-hours time, based on organizational needs. This position works in-office where working conditions, lighting, temperature, audio, and workspace are all sufficient.
PHYSICAL REQUIREMENTS
Work requires the ability to constantly operate a computer and the ability to read, type, and communicate. Work may require the ability to move work-related supplies weighing up to 10-15 pounds.
DISCLAIMER/INTENT AND FUNCTION OF JOB DESCRIPTIONS
The above information on this description has been designed to indicate the general nature and level of work performed by team members within this classif
Company & context
Evidence is labeled so you can tell internal community data from public sources.
Context may refresh in the background.
Trust-check this listing
Verify scam risk and ghost-job signals before you apply.
Related roles
Browse more remote Software Engineer jobs.
Source: We Work Remotely • Last updated Jun 24, 2026