Summary
TRM Labs provides AI-powered intelligence solutions that help public and private sector agencies investigate and disrupt crime. The Data Platform Engineer will operate and optimize the StarRocks-backed serving layer supporting government cloud investigations, build reliable data pipelines, troubleshoot production issues, and improve the resilience and compliance of distributed data infrastructure.
Responsibilities
- You will own performance tuning on the StarRocks serving layer, using AI-assisted query profiling (Claude, internal tooling) to find and fix slow query patterns before they become customer-facing incidents
- You will build and harden data pipelines feeding government cloud investigations, using AI code review workflows to ship reliable changes faster in a high-compliance environment where mistakes are costly
- You will reduce single-point-of-failure risk on GovCloud data infrastructure by becoming the second engineer who can independently operate and troubleshoot the serving layer, cutting incident response time when the primary owner is unavailable
- You will use AI-assisted debugging and log analysis to triage production issues in a regulated environment, turning multi-hour investigations into rapid root-cause fixes
Skills
- U.S. citizenship is required for this role due to government cloud data access requirements
- Hands-on experience operating distributed OLAP or serving-layer systems (StarRocks, Trino, ClickHouse, or similar), including query tuning and performance optimization at scale
- Experience owning data pipeline reliability and incident response, and comfort using AI tools (Claude, Cursor, or similar) to accelerate debugging, code review, and documentation
- Independent ownership mindset: you can pick up an unfamiliar piece of production infrastructure, use AI-assisted research and code exploration to ramp quickly, and take on-call responsibility with minimal oversight
Qualifications
Must Haves
- U.S. citizenship is required for this role due to government cloud data access requirements
- Hands-on experience operating distributed OLAP or serving-layer systems (StarRocks, Trino, ClickHouse, or similar), including query tuning and performance optimization at scale
- Experience owning data pipeline reliability and incident response, and comfort using AI tools (Claude, Cursor, or similar) to accelerate debugging, code review, and documentation
- Independent ownership mindset: you can pick up an unfamiliar piece of production infrastructure, use AI-assisted research and code exploration to ramp quickly, and take on-call responsibility with minimal oversight
Benefits