Site Reliability Engineer

Posted 3 days ago

edge search formerly alpha search advisorsNew York (NY)

SENIORITY

Lead

Apply

About the role

We are seeking a Kubernetes Engineer to join our team. In this role, you will be instrumental in shaping and executing our containerization strategy, optimizing resource utilization, and ensuring robust disaster recovery and business continuity. The ideal candidate has experience working in high performance environments and has worked on the Kubernetes internals. The Kubernetes Engineer will work with application and fellow infrastructure teams to design solutions and troubleshoot issues. Key Responsibility: This role is hands on, requiring direct interaction with platform users to understand their requirements. The ideal candidate will translate these requirements into effective solutions, then build and configure the solutions. We prioritize self-documenting, version-controlled code and configurations over traditional wiki pages and written documentation. Design and drive standards for compute usage, encompassing virtual machines and containers. Establish processes to ensure applications are host-neutral and deployable across various data center and office environments. Support firmwide disaster recovery and business continuity initiatives. Leverage data to inform strategies and decision making.
Required Qualifications: Bachelor's degree in computer science, or a related technical discipline, or an equivalent experience. Experience building and running production Kubernetes clusters Deep understanding of Linux and its network stack Experience with observability techniques including logs, metrics, traces, and profiles. Experience deploying and managing services on the Google Cloud Platform(GCP).Experience writing production-grade code in Go, Python or Rust. Experience developing with Git, issue tracking, code reviews and CI/CD pipelines.
Preferred Qualifications: Proficiency in infrastructure provisioning/management tools (e.g. Ansible, Puppet, Terraform, Packer).Experience with ArgoCD, Helm, and eBPFAdvanced knowledge of TCP/IP networking, architecture, and core technologies (such as DNS, DHCP, HTTP, Routing, VPN).Ability to manage and implement large scale infrastructure projects.

Before you apply

Applying takes about a minute. These four things decide how fast it moves after that.

Your profile is current

It's what we read first. Occupations, seniority and locations matter more than a long history.

Two examples you can talk through

Not a portfolio — just two pieces of work where you can explain the decisions and what you'd change.

A number in mind

What you're on now and what would make you move. We negotiate better when we know both.

Your notice period

Employers plan around it, and it's the question that stalls offers most often.

Once you apply, someone reads it and calls you before anything reaches the employer — usually within two working days.

More like this