Huron logo
Huron
Posted 79 days agoVerified live 8h ago

Data Engineer - AI, Agents, & Context - Clinical (Sr. Associate)

Brief overview

Remote
UndergradOr in progress
$110k–$150k/yrStated range
3+ yrsMinimum
94 H-1B approvalsDept. of Labor
9 green cardsCertified filings
SQLPythonScalaJavaCloud data pipelinesUnstructured data processingSearch and retrievalVector searchEmbeddingsSemantic modelingData governanceRBACABACPII redactionAudit logging

About the company

Huron is a global professional services firm that collaborates with clients to put possible into practice by creating sound strategies, optimizing operations, accelerating digital transformation, and empowering businesses and their people to own their future.

Visa sponsorship history

4 years sponsoring, last filed FY2026

Data powered by U.S. Department of Labor. This does not guarantee sponsorship for this specific role.
94H-1B approved
98%approval rate
25new H-1B hires
9PERM certified
$155,000median wage / yr
H-1B Petition ApprovalsVisas USCIS actually granted: the strongest sign the company sponsors.
202329
202430
202534
20261
LCA Certified ApplicationsAn early filing step, not a visa approval: it signals intent, not confirmed sponsorship.
202312
20247
20253
20261
Green Card (PERM) FilingsCertified green card filings: a long-term commitment to international hires.
20235
20241
20253
Top sponsored roles
Oracle Consulting ManagerOracle Consulting Senior ManagerWorkday Consulting ManagerDirectorDigital Consulting Director, Intelligent Automation
Sponsored employees from
IndiaCanadaPakistan

Job description

Summary

Huron is a consulting firm that helps healthcare organizations drive growth and enhance performance. The Data Engineer role focuses on building and maintaining AI data capabilities to improve clinical outcomes and optimize business operations within the healthcare sector.

Responsibilities

  • Build and contribute to the AI context platform
  • Implement end-to-end pipelines: ingestion → parsing/chunking → enrichment → embeddings → vector indexing → retrieval/serving
  • Build and maintain patterns for incremental refresh, backfills, re-embeddings, deduplication, and lineage across unstructured sources
  • Contribute to retrieval quality improvements (query strategies, hybrid search, metadata filtering) in partnership with AI engineers
  • Deliver semantic and governed data products
  • Implement semantic layers (metrics/entities) that power BI and agent reasoning consistently
  • Apply established data contracts and context contracts for AI inputs (schemas, metadata requirements, freshness, citation expectations)
  • Ensure datasets and indexes are documented and reusable
  • Support reliability and performance across assigned workstreams: monitoring, alerting, runbooks, and incident response
  • Contribute to cost and latency optimization across warehouse/lakehouse and vector infrastructure
  • Apply security-by-design patterns: RBAC/ABAC, PII redaction, retention controls, and audit logging
  • Follow established guardrails for AI access to enterprise knowledge in coordination with Security/Legal/Compliance

Skills

  • BA or BS required, preferably in Computer Science, Engineering, or a technology-based discipline
  • 3–6 years in data engineering or data platform roles with strong hands-on delivery
  • Strong SQL and Python (or Scala/Java); solid production engineering habits
  • Experience designing and operating cloud data pipelines at scale
  • Experience working with unstructured data processing and search/retrieval concepts
  • Clear communicator who can work effectively across technical and functional teams
  • Hands-on experience with vector search and embeddings (pgvector/Pinecone/Weaviate/OpenSearch/Elastic) and retrieval patterns (semantic retrieval, hybrid search, reranking)
  • Experience supporting LLM applications (RAG, agent tool interfaces, evaluation/observability)
  • Familiarity with knowledge graphs/semantic modeling or metrics layers
  • Experience in regulated environments and data governance programs

Qualifications

Must Haves

  • BA or BS required, preferably in Computer Science, Engineering, or a technology-based discipline
  • 3–6 years in data engineering or data platform roles with strong hands-on delivery
  • Strong SQL and Python (or Scala/Java); solid production engineering habits
  • Experience designing and operating cloud data pipelines at scale
  • Experience working with unstructured data processing and search/retrieval concepts
  • Clear communicator who can work effectively across technical and functional teams

Nice to Haves

  • Hands-on experience with vector search and embeddings (pgvector/Pinecone/Weaviate/OpenSearch/Elastic) and retrieval patterns (semantic retrieval, hybrid search, reranking)
  • Experience supporting LLM applications (RAG, agent tool interfaces, evaluation/observability)
  • Familiarity with knowledge graphs/semantic modeling or metrics layers
  • Experience in regulated environments and data governance programs

Benefits

  • Eligible to participate in Huron’s annual incentive compensation program, which reflects Huron’s pay for performance philosophy
  • Eligible to participate in Huron’s benefit plans which include medical, dental and vision coverage and other wellness programs

More jobs like this