Summary
Accenture Federal Services helps the US federal government strengthen national security, public safety, civilian, and military health missions through technology and ingenuity. The Data Engineer will support a client in San Antonio by building CI/CD pipelines, managing cloud environments, implementing Infrastructure as Code, and designing large-scale data pipelines for mission-critical ingestion, processing, and analytics workflows.
Responsibilities
- Building, automating, and optimizing the infrastructure and processes that enable rapid, reliable software delivery
- Working closely with developers to integrate automation, improve deployment workflows, and monitor application performance to maintain high availability across environments
- Fostering a culture of collaboration, continuous improvement, and shared ownership across engineering teams
- Proactively identifying bottlenecks, recommending architectural enhancements, and championing practices such as containerization, configuration-as-code, and automated testing
- Aligning tooling, processes, and infrastructure with business needs to help teams deliver features faster, respond to issues quickly, and maintain a stable, secure operational posture
- Designing and enhancing large-scale data pipelines leveraging Hadoop, HDFS, Apache NiFi, Apache Spark, and Accumulo to support mission-critical data ingestion, processing, and analytics workflows
- Optimizing distributed systems performance, improving data reliability, and maintaining resilient ingest and analytic patterns across the Apache ecosystem
Skills
- 3 years of experience in CI/CD pipeline development and automation
- 3 years of experience with cloud infrastructure management and monitoring
- 3 years of experience with Infrastructure as Code and configuration management
- 3 years of experience with containerization and automated testing
- Must have an active Secret clearance
- Experience with troubleshooting and continuous process improvement
- Knowledge of tuning JVM/GC, YARN, tablet servers, or NiFi performance—aligned with the types of stability and modernization work reflected in internal AFS/Recro big-data engagements
- Experience supporting cyber operations or data engineering work in environments using Accumulo, NiFi, and Spark pipelines
- Hands-on experience with distributed data platforms including Hadoop and HDFS
- Practical experience designing, maintaining, or troubleshooting data ingestion flows using Apache NiFi
- Experience building or optimizing Spark-based data processing jobs (PySpark, Scala, or Java)
- Familiarity with Accumulo as a high-ingest, low-latency data store, including integration patterns, bulk loading, or iterator development
- TS/SCI preferred
Qualifications
Must Haves
- 3 years of experience in CI/CD pipeline development and automation
- 3 years of experience with cloud infrastructure management and monitoring
- 3 years of experience with Infrastructure as Code and configuration management
- 3 years of experience with containerization and automated testing
- Must have an active Secret clearance
Nice to Haves
- Experience with troubleshooting and continuous process improvement
- Knowledge of tuning JVM/GC, YARN, tablet servers, or NiFi performance—aligned with the types of stability and modernization work reflected in internal AFS/Recro big-data engagements
- Experience supporting cyber operations or data engineering work in environments using Accumulo, NiFi, and Spark pipelines
- Hands-on experience with distributed data platforms including Hadoop and HDFS
- Practical experience designing, maintaining, or troubleshooting data ingestion flows using Apache NiFi
- Experience building or optimizing Spark-based data processing jobs (PySpark, Scala, or Java)
- Familiarity with Accumulo as a high-ingest, low-latency data store, including integration patterns, bulk loading, or iterator development
- TS/SCI preferred
Benefits
- Hands-on experience, certifications, industry training and more