Site Reliability Engineer (SRE)

Invisible AI Remote

Company

Invisible AI

Location

Remote

Type

Full Time

Job Description

At Invisible AI, we are building the future of computer vision. Today, our core focus is on developing an end-to-end platform that can digitize manufacturing operations. We deploy edge AI cameras to digitize all steps of manual assembly work which helps people-driven manufacturing be accurate, reliable, and safe. Coming from the world of self-driving cars, the founders of Invisible AI have years of experience in building and deploying large-scale AI & Machine Learning pipelines. Join us and help build a company that will deliver the endless possibilities of computer vision to real-world customers!


As a Site Reliability Engineer, you will build the technology to enable our platform to deploy, run, and monitor Invisible AI’s software at scale across tens of independent deployments and thousands of devices. The SRE works closely with all other engineering teams and owns internal tools to enable faster development and deployment, like secure ephemeral debug environments, streamlined access controls, CI/CD systems, and a custom in-house device management platform for device configuration and software releases.

Responsibilities:

  • Design, build, and maintain scalable and resilient infrastructure on the edge.
  • Develop automation and infrastructure-as-code solutions using Terraform, Ansible, and scripting languages (Python, Bash).
  • Deploy and manage containerized applications using Docker and related technologies.
  • Ensure system observability by building and optimizing monitoring systems, particularly using Prometheus.
  • Troubleshoot and optimize Linux-based systems (e.g., Red Hat, CentOS, Ubuntu).
  • Collaborate with security teams to implement robust security practices and ensure compliance with best practices.
  • Work closely with software engineers to improve system performance, reliability, and deployment pipelines.
  • Support and maintain networking infrastructure, including troubleshooting protocols and configurations.
  • Manage cloud and on-premise infrastructure, with a focus on automation and scalability.
  • Contribute to incident response, postmortems, and process improvements.

Requirements:

  • 5+ years of experience building and managing infrastructure at scale, particularly on the edge.
  • Proficiency in Python, Docker, Linux systems, and scripting (Bash, Python).Strong expertise with infrastructure automation tools (Terraform, Ansible).Experience managing observability and monitoring systems, particularly Prometheus.
  • Deep understanding of networking concepts and protocols.
  • Familiarity with cloud platforms (AWS, Azure, Google Cloud) is a plus.
  • Experience with Windows Services/VMs is a plus.
  • Excellent problem-solving skills, with attention to detail.
  • Strong communication and collaboration skills to work across teams.
  • Bachelor’s degree in Computer Science, Information Technology, or a related field, or equivalent experience.

Our compensation package plays a big part in how we value your impact on our mission. Our base pay is one part of our total compensation package and is determined within a range. This provides the opportunity to progress as you grow and develop within a role. The estimated base salary guideline range for this role is between $110,000-$170,000 and may be modified. This will vary based on various factors, including market and individual qualifications objectively assessed during the interview process. In addition to base salary, your compensation package will include additional components such as equity, sales incentive pay (for sales roles), and benefits. Invisible AI is an equal-opportunity employer. We do not discriminate based on age, ethnicity, gender, nationality, religious belief, or sexual orientation.

Apply Now

Date Posted

01/21/2025

Views

0

Back to Job Listings ❤️Add To Job List Company Info View Company Reviews
Positive
Subjectivity Score: 0.9

Similar Jobs

Linux Support Engineer - Voltage Park

Views in the last 30 days - 0

Voltage Park is seeking a Linux Support Engineer for a fulltime remote position The ideal candidate will have command line level Linux sys administrat...

View Details

Technical Architect - CDW

Views in the last 30 days - 0

CDW offers a rewarding career opportunity for a Technical Architect with expertise in ServiceNow The role involves delighting customers by collaborati...

View Details

Federal Security Solutions Engineer - Rapid7

Views in the last 30 days - 0

Rapid7 is seeking a Federal Solutions Engineer with 5 years of experience in cybersecurity solutions engineering or technical sales focusing on federa...

View Details

Sales Engineer - Dandy

Views in the last 30 days - 0

Dandy a venturebacked company is revolutionizing the 200B dental industry with advanced technology They are looking for a Sales Engineer with 5 years ...

View Details

Engineering Manager (Group Practice Tooling & Provider CX) - Headway

Views in the last 30 days - 0

Headway is a mental healthcare company founded in 2019 aiming to build a new mental health care system accessible to everyone They have a national net...

View Details

Engineering Manager (Claims Platform) - Headway

Views in the last 30 days - 0

Headway is a mental healthcare company founded in 2019 aiming to build a new mental health care system accessible to everyone They have a national net...

View Details