Site Reliability Engineer (6266)

1 week, 1 day ago
Contract
Junior
DevOps and Infrastructure
Dan.com - a GoDaddy brand

Dan.com - a GoDaddy brand

Dan.com is a GoDaddy brand that offers a wide range of domain-related products and services. With a focus on simplifying the domain buying and selling process, Dan.com provides a user-friendly platform for individuals and businesses to search, register...

Internet Software & Services

Description

  • Develop and maintain infrastructure automation solutions using Ansible for cloud environments.
  • Design, implement, and enhance CI/CD pipelines, testing frameworks, and operational tooling.
  • Troubleshoot complex Linux-based infrastructure and distributed systems issues.
  • Build automation for rapid, repeatable deployment of regional, sovereign, and purpose-built cloud environments.
  • Collaborate with engineering teams, product management, and cross-functional stakeholders on operational improvements.
  • Monitor infrastructure performance and implement enhancements to reduce overhead and improve scalability.
  • Contribute to automation and engineering best practices for reliable cloud platform operations.
  • Participate in internal practice meetings, thought leadership, case studies, and networking events.
  • Work with leadership on career fast-track opportunities and internal practice development.

Requirements

  • 2+ years of experience in Site Reliability Engineering, DevOps, Infrastructure Engineering, or a related role.
  • Experience developing and maintaining infrastructure automation using Ansible.
  • Experience programming in Ruby and writing automated tests using RSpec or similar frameworks.
  • Experience administering and troubleshooting Linux-based systems and distributed infrastructure environments.
  • Experience designing, implementing, and maintaining CI/CD pipelines, including GitLab CI.
  • Experience supporting large-scale infrastructure environments with hundreds or thousands of systems.
  • Must be eligible to work on FedRAMP projects.
  • Must be a U.S. citizen working from U.S. soil.
  • Bachelor's degree in a relevant field or equivalent work experience.
  • Experience with AWS or other public cloud platforms and hybrid infrastructure environments (preferred).
  • Knowledge of monitoring, observability, and site reliability engineering practices and tools (preferred).
  • Familiarity with Kubernetes concepts and containerized application platforms (preferred).
  • Experience using AI-assisted development tools to improve development and operational productivity (preferred).

Benefits

  • 100% remote work within the United States.
  • 24-month duration.
  • Comprehensive medical benefits.
  • Dental, vision, and life insurance.
  • 401(k) plan with matching.
  • Paid holidays.
  • Networking, career learning, and development programs.

Interested in this position?

Apply directly on the company website

Apply Now

Similar Roles

Network Reliability Engineer

Margo Bank Professional Services

Network Reliability Engineer at Warsaw Consulting – Polska Team, working remotely to build and operate AI infrastructure with a focus on monitoring, incident response, and service reliability.

Ansible Bash CI/CD Debian DNS Elasticsearch GitLab Go Grafana Linux Load Balancing MariaDB Prometheus Python SaltStack TCP/IP Ubuntu
1 month, 4 weeks ago

Reliability Engineer

Sapsol Technologies 51-250 Internet Software & Services

A medical device company is seeking a Reliability Engineer to develop and execute reliability requirements, testing, and risk analyses for IVD products in a highly regulated environment.

MATLAB Python R
2 months ago

Blockchain Site Reliability Engineer

InfStones 51-250 Internet Software & Services

InfStones is hiring a remote Blockchain Site Reliability Engineer in Dallas to ensure the reliability, availability, and performance of its blockchain node infrastructure.

Docker Ethereum Go Grafana JavaScript Kubernetes Linux Prometheus Python Rust Solana
3 months, 4 weeks ago

Senior Site Reliability Engineer

Intuition Machines 51-250 Life Sciences Tools & Services

Intuition Machines is hiring a Senior Site Reliability Engineer to support its internet-scale AI/ML security products, with a focus on improving the performance, availability, security, and cost efficiency of systems serving millions of users.

C++ CI/CD Cloudflare Cybersecurity Go JavaScript Kubernetes Load Balancing Machine Learning Python Rust
3 months, 4 weeks ago

You're on a roll! Sign up now to keep applying.

Sign Up

Already have an account? Log in

Used by 14,729+ remote workers