Role Description
We are seeking a highly skilled DevOps/Cloud Platforms Engineer to join our engineering organization and help drive the next generation of our infrastructure and platform capabilities. In this role, you will be responsible for:
-
Designing, operating, and optimizing cloud environments.
-
Maintaining cloud infrastructure.
-
Creating internal developer platforms that support fast, reliable, and secure software delivery.
-
Working closely with engineering, QA, data, and product teams to ensure our systems are scalable, resilient, and easy to use.
The ideal candidate brings a strong blend of cloud infrastructure expertise, DevOps mindset, and hands-on development skills. You will:
-
Automate wherever possible.
-
Use Infrastructure as Code to maintain consistency.
-
Leverage modern DevOps tooling, including AI-assisted automation to enhance operational efficiency and developer productivity.
This is an operations role for someone who thrives in solving complex technical problems, improving processes, and building platforms that accelerate engineering velocity.
Qualifications
-
4-8+ years of experience in DevOps, Cloud Engineering, Systems Administration, or similar infrastructure-focused roles.
-
Familiar with GIT and comfortable with at least one systems programming language (Golang) and one scripting language (Python, Bash).
-
Proficiency with Infrastructure as Code (Terraform, Pulumi, or CloudFormation) and configuration management (Ansible, Chef, or SaltStack).
-
Working knowledge of at least one of observability stacks (Prometheus, Grafana, ELK/OpenSearch, Datadog, etc.) and operational troubleshooting.
-
Strong hands-on experience with Kubernetes management in production environments.
-
Experience designing, developing, and/or troubleshooting distributed systems.
-
Comfort with shells on *nix family systems.
-
B.S. degree or equivalent experience in Engineering, Computer Science or a related field.
Requirements
-
Own the deployment, maintenance, and lifecycle management of systems supporting engineering (Kubernetes clusters, container registries, artifact systems, and internal developer platforms).
-
Troubleshoot complex infrastructure and application issues, driving root-cause analysis and developing long-term remediation solutions.
-
Design, build, and maintain cloud infrastructure across major cloud providers (AWS, GCP, Azure).
-
Develop, enhance, and maintain CI/CD pipelines using modern DevOps tooling (GitHub Actions, ArgoCD, Terraform, etc.).
-
Develop internal tooling and automation using Terraform, Python, Go, or similar languages to streamline operational tasks and improve developer productivity.
-
Implement and manage security best practices across cloud environments, including identity management, secrets handling, audit logging, and network controls.
-
Leverage AI/ML tools to automate repetitive DevOps tasks and operational workflows.
Preferred Qualifications
-
Certifications from cloud service providers (e.g., AWS DevOps, GCP DevOps Engineer).
-
Strong software development skills, with experience building production-grade services, automation, APIs, and infrastructure tooling.
-
Proficiency in Go, Python, or other modern programming languages.
-
Strong cloud networking fundamentals, including VPC/VNet, subnets, NAT Gateways, Transit Gateways, routing, DNS, load balancing, and private connectivity.
-
Hands-on experience with VPN and secure networking technologies such as Tailscale, Headscale, WireGuard, or similar solutions.
-
Strong Kubernetes experience, including cluster architecture, networking, service discovery, ingress, observability, and troubleshooting.
-
Understanding of distributed systems concepts and experience working with highly available, scalable, and fault-tolerant systems.
-
Natural curiosity and a strong drive to learn new and adjacent technologies.
-
Experience applying AI/ML tools (e.g., GitHub Copilot, cloud AI services, LLM-based automation, anomaly detection tools) to streamline DevOps and infrastructure workflows is a strong plus.
-
Experience with LangChain, AI agents, and MCP (Model Context Protocol), including building and integrating agent-based solutions.
-
Hands-on experience implementing AI-enhanced threat detection and security systems.
Benefits
-
SingleStore delivers our cloud-native database with the speed and scale to power the worldβs data-intensive applications.
-
Salary is based on permissible, non-discriminatory factors such as skills, experience, and geographic location.
-
Roles in a variety of locations across the United States.
-
Additional rewards, including merit increases and annual bonuses for certain roles.