Intuition Machines

Intuition Machines

Intuition Machines is a leading company in the field of Privacy Preserving AI/ML. They specialize in turning AI/ML research into platforms and services that prioritize user privacy. Their products, including hCaptcha.com, are widely used and have a sig...

Life Sciences Tools & Services
51-250

Description

  • Work with large-scale systems handling millions of requests per second across multiple cloud providers.
  • Develop solutions that improve performance, availability, security, and cost-effectiveness.
  • Maintain system uptime and speed while keeping development teams productive.
  • Ensure peer releases improve quality, security, uptime, delivery speed, threat detection, and customer engagement.
  • Source improvement ideas and priorities from customers, internal teams, and system metrics.
  • Make rapid decisions to support a small-team, fast-iteration environment.
  • Work across infrastructure, data, and application logic layers to build system-level solutions.
  • Collaborate with customer teams and internal stakeholders in a flat organization.

Requirements

  • Expert-level experience with Kubernetes.
  • Expert-level experience monitoring applications, infrastructure, and networks.
  • Software engineering background with backend development experience in Kubernetes-based systems.
  • Strong programming skills in Python, JavaScript, Go, C++, or Rust.
  • Strong understanding of networking, proxies, and content delivery networks, including Cloudflare.
  • Experience with multi-cloud environments, including virtual networking, load balancing, and web application firewalls.
  • Strong experience with CI/CD.
  • Hands-on experience in high-scale, high-uptime, and high-reliability environments.
  • Minimum of 6 years of hands-on experience in engineering, DevOps, or SRE roles.
  • Familiarity with distributed systems, including queue-first architectures and sharding.
  • Demonstrated ability to gather requirements, solve problems, and make recommendations.
  • Preferred: familiarity with security frameworks, attack vectors, botnets, and impact analysis.
  • Must pass pre-employment screening, including third-party verification of work history, education, and identity, plus a final in-person interview and identity verification in the country of residence.

Benefits

  • Fully remote position with flexible working hours.
  • A global team of colleagues distributed around the world.
  • Modern development and deployment workflows with an emphasis on shipping early and often.
  • High-impact work with lots of users, happy customers, and high growth.
  • Direct interaction with customer teams in a flat organization.
  • Commitment to equality of opportunity and an inclusive work environment.

Interested in this position?

Apply directly on the company website

Apply Now

Similar Roles

Lead Site Reliability Engineer (Performance & Scalability) | Contract | Remote US

Tech Holding 51-250 Internet Software & Services

Tech Holding is hiring a contract Lead Site Reliability Engineer to assess and improve the performance, reliability, and scalability of its platform for current and future demand.

CI/CD
1 week, 6 days ago

Reliability Engineer

Sapsol Technologies 51-250 Internet Software & Services

A medical device company is seeking a Reliability Engineer to develop and execute reliability requirements, testing, and risk analyses for IVD products in a highly regulated environment.

MATLAB Python R
2 months, 3 weeks ago

Blockchain Site Reliability Engineer

InfStones 51-250 Internet Software & Services

InfStones is hiring a remote Blockchain Site Reliability Engineer in Dallas to ensure the reliability, availability, and performance of its blockchain node infrastructure.

Docker Ethereum Go Grafana JavaScript Kubernetes Linux Prometheus Python Rust Solana
4 months, 2 weeks ago

DevOps Engineer (Cloud) - Freelance

Lingaro 5K-10K IT Services

Build and evolve an Azure-based monitoring and observability platform that improves system reliability, data quality, operational insight, automation, and cloud cost management across supply chain systems.

Ansible Azure Bash CI/CD Databricks Docker GitHub Actions Grafana Kafka Kubernetes Linux Power BI PowerShell Prometheus Python SonarQube Terraform Windows Server
9 hours, 34 minutes ago

You're on a roll! Sign up now to keep applying.

Sign Up

Already have an account? Log in

Used by 14,729+ remote workers