Summary
Procom is recruiting on behalf of a technology client for a Backend Engineer, AI. The role owns the inference and orchestration layer for AI interactions, building and operating production systems and APIs while optimizing performance, monitoring, and reliability.
Responsibilities
- Build and operate backend systems that serve AI-powered features in production
- Design inference pipelines, orchestration layers, and service boundaries around models
- Own production concerns: monitoring, logging, alerting, and incident response
- Optimize latency and throughput across inference, caching, batching, and streaming
Skills
- Strong backend engineering fundamentals in production environments
- Experience running high-throughput, low-latency services
- Familiarity with AI inference patterns (LLMs, embeddings, multimodal)
- Comfortable debugging distributed systems under load
- Bias toward shipping and learning from production behavior
- This position is a remote position based in Austin, Texas
- Remote position based in Austin, Texas, United States
- Experience with Python and NodeJs
- Familiarity with Pytorch and open-source LLMs
- Knowledge of SQL & noSQL databases
- Experience with Kubernetes and Docker
- Strong communication and teamwork skills
Qualifications
Must Haves
- Strong backend engineering fundamentals in production environments
- Experience running high-throughput, low-latency services
- Familiarity with AI inference patterns (LLMs, embeddings, multimodal)
- Comfortable debugging distributed systems under load
- Bias toward shipping and learning from production behavior
- This position is a remote position based in Austin, Texas
- Remote position based in Austin, Texas, United States
Nice to Haves
- Experience with Python and NodeJs
- Familiarity with Pytorch and open-source LLMs
- Knowledge of SQL & noSQL databases
- Experience with Kubernetes and Docker
- Strong communication and teamwork skills
Benefits
- Permanent full-time position
- Remote position based in Austin, Texas, United States