Intuition Machines

Intuition Machines

Intuition Machines is a leading company in the field of Privacy Preserving AI/ML. They specialize in turning AI/ML research into platforms and services that prioritize user privacy. Their products, including hCaptcha.com, are widely used and have a sig...

Life Sciences Tools & Services
51-250

Description

  • Work with large-scale systems handling millions of requests per second across multiple cloud providers.
  • Develop solutions that improve performance, availability, security, and cost-effectiveness.
  • Maintain system uptime and speed while keeping development teams productive.
  • Ensure peer releases improve quality, security, uptime, delivery speed, threat detection, and customer engagement.
  • Source improvement ideas and priorities from customers, internal teams, and system metrics.
  • Make rapid decisions to support a small-team, fast-iteration environment.
  • Work across infrastructure, data, and application logic layers to build system-level solutions.
  • Collaborate with customer teams and internal stakeholders in a flat organization.

Requirements

  • Expert-level experience with Kubernetes.
  • Expert-level experience monitoring applications, infrastructure, and networks.
  • Software engineering background with backend development experience in Kubernetes-based systems.
  • Strong programming skills in Python, JavaScript, Go, C++, or Rust.
  • Strong understanding of networking, proxies, and content delivery networks, including Cloudflare.
  • Experience with multi-cloud environments, including virtual networking, load balancing, and web application firewalls.
  • Strong experience with CI/CD.
  • Hands-on experience in high-scale, high-uptime, and high-reliability environments.
  • Minimum of 6 years of hands-on experience in engineering, DevOps, or SRE roles.
  • Familiarity with distributed systems, including queue-first architectures and sharding.
  • Demonstrated ability to gather requirements, solve problems, and make recommendations.
  • Preferred: familiarity with security frameworks, attack vectors, botnets, and impact analysis.
  • Must pass pre-employment screening, including third-party verification of work history, education, and identity, plus a final in-person interview and identity verification in the country of residence.

Benefits

  • Fully remote position with flexible working hours.
  • A global team of colleagues distributed around the world.
  • Modern development and deployment workflows with an emphasis on shipping early and often.
  • High-impact work with lots of users, happy customers, and high growth.
  • Direct interaction with customer teams in a flat organization.
  • Commitment to equality of opportunity and an inclusive work environment.

Interested in this position?

Apply directly on the company website

Apply Now

Similar Roles

SRE Engineer Contractor

Beauty For All Industries (BFA) 51-200 information technology & services

IPSY is seeking a remote Site Reliability Engineer Contractor in Mexico or Colombia, covering the PST time zone, to improve the availability, resilience, and operational reliability of its beauty membership platform.

Amplitude AWS Bash CI/CD Contentful Datadog Grafana Microservices Netlify New Relic OpsGenie PagerDuty Prometheus Python Terraform
1 week, 5 days ago

Site Reliability Engineer, Tech Lead

Loadsmart 251-1K Air Freight & Logistics

Loadsmart is hiring a remote SRE Tech Lead in Brazil to build and operate its internal engineering platform, improve reliability, and enable safe, dependable applications across engineering teams.

Ansible AWS Bash Chef CI/CD Docker Kubernetes PostgreSQL Python Terraform
3 weeks, 2 days ago

Incident management / reliability / SRE Evaluator

Weekday 11-50 Construction & Engineering

An independent contractor Evaluator will remotely assess AI-generated documents, spreadsheets, and presentations for accuracy, rigor, and quality using incident management, reliability, and SRE expertise.

3 weeks, 3 days ago

Network Reliability Engineer

Margo Bank Professional Services

Network Reliability Engineer at Warsaw Consulting – Polska Team, working remotely to build and operate AI infrastructure with a focus on monitoring, incident response, and service reliability.

Ansible Bash CI/CD Debian DNS Elasticsearch GitLab Go Grafana Linux Load Balancing MariaDB Prometheus Python SaltStack TCP/IP Ubuntu
3 months, 4 weeks ago

You're on a roll! Sign up now to keep applying.

Sign Up

Already have an account? Log in

Used by 14,729+ remote workers