Summary
Cube is a company developing infrastructure for data, analytics, and AI-agent applications. The Platform Engineer will design, build, and maintain the core platform powering Cube Cloud and Cube Core, focusing on reliability, scalability, performance, and developer productivity. The role includes operating cloud-native and Kubernetes-based infrastructure, scaling multi-tenant systems, improving observability and disaster recovery, and supporting deployment and incident response practices.
Responsibilities
- Designing and maintaining cloud-native infrastructure for a globally distributed analytics platform
- Building internal platform tooling to improve developer velocity, observability, and deployment safety
- Scaling multi-tenant systems that serve large volumes of analytical queries with strict latency requirements
- Improving reliability, fault tolerance, and disaster recovery for Cube cloud platform
- Operating Kubernetes-based environments and evolving our deployment and release pipelines
- Implementing monitoring, alerting, and incident response practices for a high-availability system
Skills
- Strong experience with cloud infrastructure (AWS, GCP, or similar)
- Hands-on experience with Kubernetes, container orchestration, and infrastructure-as-code (Terraform or equivalent)
- Solid understanding of distributed systems, networking, and Linux internals
- Experience building and operating production systems with high availability and performance requirements
- Proficiency in at least one programming language used for platform development (Go, Rust, TypeScript, Python, or similar)
- Familiarity with CI/CD systems and release automation
- Experience with observability stacks (Prometheus, Grafana, OpenTelemetry, etc.)
- Good communication skills and ability to work effectively in a remote, async environment
- Fluent English
- Background in security best practices for cloud platforms
- Experience operating data-intensive or analytics-heavy platforms
- Hands-on use of AI coding tools (Claude Code, Cursor, Copilot, or similar) in your day-to-day engineering workflow
- Contributions to open-source infrastructure or platform projects
Qualifications
Must Haves
- Strong experience with cloud infrastructure (AWS, GCP, or similar)
- Hands-on experience with Kubernetes, container orchestration, and infrastructure-as-code (Terraform or equivalent)
- Solid understanding of distributed systems, networking, and Linux internals
- Experience building and operating production systems with high availability and performance requirements
- Proficiency in at least one programming language used for platform development (Go, Rust, TypeScript, Python, or similar)
- Familiarity with CI/CD systems and release automation
- Experience with observability stacks (Prometheus, Grafana, OpenTelemetry, etc.)
- Good communication skills and ability to work effectively in a remote, async environment
- Fluent English
Nice to Haves
- Background in security best practices for cloud platforms
- Experience operating data-intensive or analytics-heavy platforms
- Hands-on use of AI coding tools (Claude Code, Cursor, Copilot, or similar) in your day-to-day engineering workflow
- Contributions to open-source infrastructure or platform projects
Benefits
- Fully remote work; you can work from anywhere
- Highly collaborative team