Senior Software Engineer (SRE)
Posted 12 days ago
emedusDoral (FL)
Software DevelopersComputer Systems Design Services
SENIORITY
Lead
About the role
As a Senior Software Engineer in SRE at eMed, you will play a key role in ensuring our platform is highly available, secure, and performant. You’ll lead reliability engineering efforts across production systems, drive operational excellence, and collaborate closely with application and infrastructure teams to design resilient services. This role suits an engineer with a software mindset and deep operational experience, who thrives on improving systems through automation and proactive engineering.
2.2 WHAT YOU WILL WORK ON Design and implement robust monitoring, alerting, and observability systems across all services and infrastructure
Lead reliability reviews, incident response, and post-incident analysis—focusing on prevention, learning, and long-term improvements
Improve service scalability, fault tolerance, and performance through architectural input and systems optimisation
Build and maintain automation for infrastructure management using Terraform, and delivery pipelines using GitHub Actions
Partner with software engineers to improve the operational readiness and resilience of services, including capacity planning and runbooks
Lead initiatives to reduce operational toil through tooling, automation, and process improvement
Manage and optimise our production Kubernetes and AWS environments with a focus on reliability, security, and cost-effectiveness
Contribute to security hardening efforts, including network controls, secrets management, and compliance readiness
Participate in and lead in-person stand-ups, incident reviews, and cross-team planning sessions
Share knowledge and mentor engineers on best practices in observability, incident response, and operational engineering
2.3 WHAT WE’RE LOOKING FOR: Technical Skills (Essential) Strong experience operating Kubernetes and cloud-native infrastructure (preferably EKS on AWS) in production environments
Proficiency in AWS services, including networking, compute, IAM, and logging/monitoring tools (e.g. CloudWatch, ELB, VPC)
Skilled in Terraform and Infrastructure as Code practices
Deep understanding of observability tooling (metrics, logs, tracing) and incident management workflows
Strong coding skills for building tools, scripts, and automation
Ability to troubleshoot complex infrastructure issues and lead delivery of reliable cloud solutions
Preferred Experience implementing SLAs, SLOs, and error budgets to guide operational priorities
Background in healthcare or other regulated industries with security and compliance requirements
Previous involvement in platform security reviews
2.4 PERKS AT WORK: Retirement Plan (401k with Company Match)
Life Insurance (Basic, Voluntary & AD&D)
Paid Time Off
Short Term & Long Term Disability
Training & Development
Catered Breakfast and Lunch 5 days a Week
#J-18808-Ljbffr
Before you apply
Applying takes about a minute. These four things decide how fast it moves after that.
Your profile is current
It's what we read first. Occupations, seniority and locations matter more than a long history.
Two examples you can talk through
Not a portfolio — just two pieces of work where you can explain the decisions and what you'd change.
A number in mind
What you're on now and what would make you move. We negotiate better when we know both.
Your notice period
Employers plan around it, and it's the question that stalls offers most often.
Once you apply, someone reads it and calls you before anything reaches the employer — usually within two working days.
More like this
