Site Reliability Engineer - Observability Job at Second Front Systems, Remote

SmpQaEVxWW16M1FOK0twNGNlR0hua0RnWXc9PQ==
  • Second Front Systems
  • Remote

Job Description

ABOUT THE ROLE

Second Front Systems' (2F) Product team is seeking a highly skilled and motivated Senior Site Reliability Engineer to join our Observability team. We are a small team working to accelerate the deployment of emerging technology into national security use-cases. We are seeking technical professionals who want to operate on the front lines of an exciting and disruptive mission.

As a Senior SRE for Second Front Systems, you'll be responsible for deploying, maintaining, and scaling our observability infrastructure across multiple DoD networks. You'll work with Kubernetes-based platforms, BigBang charts from DoD Platform One, and build automation to make our monitoring stack easier to deploy for new customers. You'll be empowered to collaborate with others to implement infrastructure that delivers unique capabilities for our commercial and government customers, including the Department of Defense.

The Observability team is looking for a strong SRE with deep DevSecOps and Kubernetes experience. Someone who has deployed and maintained monitoring infrastructure at scale, with an eye for security in highly-regulated environments. Experience with DoD software deployments, Platform One, and single-tenant architectures is highly valued.

We are a fast-growing entrepreneurial team working at the convergence of technology and national security. If this type of effort interests you, come join us!

Note: This position requires U.S. citizenship due to government contract requirements.

\n

What You’ll Do
  • Deploy and maintain observability stack (Grafana, Mimir, Prometheus) across multiple customer clusters and DoD networks
  • Build Helm chart abstractions and automation to streamline monitoring deployments for new customers
  • Troubleshoot and debug complex Kubernetes issues, networking problems, and monitoring stack failures
  • Configure and maintain BigBang charts and DoD Platform One integrations
  • Design and implement infrastructure automation using tools like Pulumi, ArgoCD, and Flux
  • Work with Istio service mesh and Keycloak for authentication in secure environments
  • Monitor and optimize performance of monitoring infrastructure across multiple environments
  • Collaborate with security teams to ensure compliance with NIST requirements and DoD standards
  • Participate in on-call rotation and incident response for production environments

Skills You’ll Bring to Our Team
  • 5+ years of Site Reliability Engineering or DevOps experience
  • Deep experience with Kubernetes administration, troubleshooting, and scaling
  • Hands-on experience deploying and maintaining observability tools (Prometheus, Grafana, Mimir/Cortex)
  • Strong understanding of Helm charts, GitOps practices, and CNCF tooling
  • Experience with service mesh technologies (Istio preferred)
  • Proven ability to debug complex distributed systems and networking issues
  • Understanding of authentication systems and security in regulated environments
  • Ability to work independently and collaborate with team members in a remote environment

Preferred Qualifications
  • Active security clearance or ability to obtain a Secret-level security clearance
  • Previous experience with DoD software deployments and Platform One
  • Experience with BigBang charts and Iron Bank containers
  • Experience working in national security or highly regulated environments
  • Familiarity with compliance frameworks (NIST, FedRAMP, etc.)
  • Experience with infrastructure as code (Pulumi, Terraform)

Technologies we Use
  • Observability : Grafana stack, Prometheus, custom alerting tools
  • Kubernetes : Helm, ArgoCD, Flux, Tekton, BigBang charts
  • Security : Istio, Keycloak, Kyverno
  • Infrastructure : AWS/GCP/Azure, Pulumi, Git/GitLab
  • Languages : YAML, Bash, Go

\n

$160,000 - $180,000 a year

Perks & Benefits

This role is full time. As a public benefit corporation, we’re a team of purpose-driven trailblazers transforming the future of U.S. national security. We hire the best to do their best and, as such, we are committed to providing the perks and benefits you need to be successful—both in- and outside the workplace.

We offer you:

Competitive Salary

100% Healthcare, vision and dental coverage

401(k) + 3% company contribution

Wellness perks (Fitness classes, mental health resources)

Equity incentive plan

Tech + office supplies stipend

Annual professional development stipend

Flexible paid time off + federal holidays off

Parental leave

Work from anywhere

Referral Bonus

Visit our careers page to learn more.

#LI-Remote

\n

Job Tags

Remote job, Full time, Contract work, Work at office, Flexible hours,

Similar Jobs

Hilton

Lead Pastry Cook - Canopy by Hilton Sioux Falls Downtown Job at Hilton

Canopy by Hilton Sioux Falls is seeking a creative and innovative Lead Pastry Cook to support the hotels culinary offerings. The Canopy at Sioux Falls is proud to be recognized as an AAA Four Diamond property, offering refined, stylish accommodations and upscale service...

ACS Air Conditioning Specialist Inc

Plumbing Senior Service Technician 5 Job at ACS Air Conditioning Specialist Inc

 ...Job Description Job Description Job description Plumbing Service Technician - Greensboro, Milledgeville, Covington, Lake Oconee areas We are seeking experienced Service Plumbing Technicians to join our teams in the Lake Oconee, GA & Covington, GA areas! These... 

Domino's Franchise

Delivery Driver(07919) - 1301C Palmetto Avenue Job at Domino's Franchise

 ...all equipment Stock ingredients from delivery area to storage, work area, walk-in cooler...  ...the following applies to team members in driver or store management positions. Job Duties...  ...product, driving and couponing. SENSING: Far vision and night vision for driving.... 

Formation Bio

Senior Data Engineer II - Electronic Health Records (EHR) Job at Formation Bio

 ...ultimately helping to bring new medicines to patients. The company is backed by investors across pharma and tech, including a16z, Sequoia, Sanofi, Thrive Capital, Sam Altman, John Doerr, Spark Capital, SV Angel Growth, and others. You can read more at the following links:... 

EQT Real Estate

Acquisition Associate Job at EQT Real Estate

 ...Estate (formerly EQT Exeter) is among the largest real estate investment managers in the world, focused on acquiring, developing and managing...  ...experience within the real estate industry. Investment banking backgrounds will also be considered. Significant attention...