Job Overview
The SRE Compliance & Security Initiatives team is seeking a Jr. Site Reliability Engineer to support the stability, scalability, and security of our FedRAMP cloud platform. Operating with a high degree of autonomy, you will focus on vulnerability management, automating security processes, and helping our infrastructure adapt to evolving regulatory requirements.
Key Responsibilities
- Vulnerability Management: Own findings from intake through validated remediation and closure. Triage and prioritize host, OS-package, container, base-image, and application-dependency findings based on exploitability, asset criticality, and deadlines.
- Remediation & Analysis: Identify the true source of vulnerabilities and coordinate fixes (dependency updates, image rebuilds, host changes, deviations, or false-positive corrections) using Ansible, GitLab CI, package managers, and container builds.
- Validation: Verify fixes using package versions, build artifacts, image manifests, deployed-host evidence, and scanner rescans before closing tickets.
- Automation & Tooling: Develop and maintain automation solutions using Ansible to improve infrastructure reliability, scalability, and security compliance.
- CI/CD & Deployment: Design and enhance deployment pipelines, testing frameworks, and operational tooling to support rapid scaling across global environments.
- Documentation & Tracking: Maintain Jira evidence, ownership escalations, backlog metrics, runbooks, and knowledge transfer to streamline processes.
Minimum Qualifications
- 2+ years of experience in Site Reliability Engineering, DevOps, Infrastructure Engineering, or a related role supporting cloud-based production environments.
- Practical vulnerability-management experience using tools like Qualys, JFrog Xray, SCA/SBOM, or equivalent.
- Working knowledge of dependency management for Ruby/Bundler, Python/pip, and Go modules (including direct vs. transitive dependencies, version constraints, lockfiles, go.mod/go.sum).
- Experience developing and maintaining infrastructure automation using Ansible.
- Strong experience administering and troubleshooting Linux-based systems and distributed infrastructure environments.
- Experience designing, implementing, and maintaining CI/CD pipelines, including GitLab CI.
- Experience supporting large-scale infrastructure environments consisting of hundreds or thousands of systems.
- Availability to be online daily from 9:00 AM to 4:00 PM PST.
Preferred Qualifications
- Familiarity with AWS or other public cloud platforms and hybrid infrastructure environments.
- Knowledge of monitoring, observability, and reliability engineering practices.
- Familiarity with Kubernetes and containerized application platforms.
- Experience leveraging AI-assisted development tools to improve automation and engineering productivity.
- Experience managing fleet-wide software deployments and providing priority incident triage.