Resilient Co

Resilient Co

Resilient Co is a technology consulting company that empowers businesses with smart solutions and diverse teams, offering resilient support in the dynamic tech industry.

Professional Services
11-50
Founded 2020

Description

  • Design, deploy, and maintain production Kubernetes clusters and related services.
  • Build and maintain Python-based automation and operational tooling.
  • Integrate and operate Prometheus for monitoring, alerting, and observability.
  • Deploy and manage Ceph storage solutions for distributed workloads.
  • Support platform modernization initiatives and migrate services to cloud-native patterns.
  • Troubleshoot and resolve issues across compute, storage, and network layers in distributed systems.
  • Collaborate with development, SRE, and operations teams to define platform requirements and SLAs.
  • Document platform designs, runbooks, and operational procedures.
  • Participate in on-call rotations and incident response to maintain platform availability.

Requirements

  • 5+ years of experience in platform, infrastructure, or site reliability engineering roles.
  • Proven experience deploying and operating Kubernetes in production.
  • Strong Python skills for automation, tooling, and operational scripts.
  • Experience implementing and operating Prometheus-based monitoring and alerting.
  • Hands-on experience with Ceph or similar distributed storage systems.
  • Cloud experience with AWS and Azure, including designing, deploying, and operating services.
  • Demonstrated ability to troubleshoot distributed systems and resolve production incidents.
  • Experience collaborating across teams to deliver platform improvements and migrations.
  • Experience with OpenSearch (preferred).
  • Proficiency with Bash scripting (preferred).
  • Familiarity with Java-based services (preferred).
  • Experience with Fluent Bit for log collection (preferred).
  • Experience working with PostgreSQL (preferred).

Benefits

  • 12-month or longer engagement.
  • PST working hours (8:00 AM - 5:00 PM).
  • Client USA holiday calendar is mandatory.
  • No overtime required.
  • BYOD laptop policy.

Interested in this position?

Apply directly on the company website

Apply Now

Similar Roles

Cloud Infrastructure Engineer (Remote LATAM)

Atmosera 51-250 IT Services

Atmosera is seeking a senior Azure Cloud Infrastructure Engineer contractor to design, implement, migrate, and operate secure cloud environments while advising clients and supporting internal operations.

Agile Ansible Azure PowerShell SQL Server Terraform Windows Server
2 days, 14 hours ago

Sr. Infrastructure Engineer, DevOps

Very Good Security 51-250 Internet Software & Services

VGS is hiring a Senior Infrastructure Engineer (DevOps) to build and operate secure, resilient cloud platforms, CI/CD systems, and internal developer tooling for its global payments infrastructure.

Argo CD AWS Bash CI/CD Docker Flux GitHub Actions GitOps Go Grafana Java Kafka Kubernetes OpenTelemetry Prometheus Python TypeScript
2 weeks, 1 day ago

Sr. Infrastructure Engineer

Very Good Security 51-250 Internet Software & Services

VGS is hiring a Senior Infrastructure Engineer to build and operate resilient, multi-region AWS infrastructure supporting high-throughput global payment systems.

Argo CD AWS Bash CI/CD Docker Flux GitHub Actions GitOps Go Grafana Java Kafka Kubernetes Load Balancing OpenTelemetry Prometheus Python Terraform TypeScript
2 weeks, 1 day ago

You're on a roll! Sign up now to keep applying.

Sign Up

Already have an account? Log in

Used by 14,729+ remote workers