Summary
Versor Investments is a pioneer in applying AI and alternative data to global equity markets, headquartered in New York. They are seeking a Data Engineer to join their Portfolio Analytics team, focusing on designing and developing a large-scale data lake while collaborating with senior researchers.
Responsibilities
- Design architecture for a data platform
- Design data pipelines based on business and functional requirements
- Extract, transform, and load logic to automate data collection and manage data processes/pipelines. This includes data quality and monitoring
- Develop data access tools to allow researchers to access data seamlessly
- Develop integration tools and analytical reports for the databases and data warehouse
- Write and review technical documents. This includes requirements and design documents for existing and future data systems, as well as data standards and policies
- Collaborate with analysts, support/system engineers, and business stakeholders to ensure data infrastructure meets constantly evolving requirements
Skills
- E., B.Tech., M.Tech., or M.Sc. in Computer Science, Computer Engineering or similar discipline from a top tier institute
- 2+ years direct experience working as a data engineer
- Experience in design, architecture and implementation of data lake, data pipelines and flows
- Experience with developing software code and APIs in one or more languages such as Python and C#
- Experience designing and deploying large scale distributed data processing systems with one or more technologies such as MS SQL Server, PostgreSQL, MongoDB, Cassandra, Redis, Hadoop, Spark, Hive, Teradata, or MicroStrategy
- A high-level understanding of automation in a cloud environment (AWS experience preferred)
- Excellent communication, presentation, and problem-solving skills
Qualifications
Must Haves
- E., B.Tech., M.Tech., or M.Sc. in Computer Science, Computer Engineering or similar discipline from a top tier institute
- 2+ years direct experience working as a data engineer
- Experience in design, architecture and implementation of data lake, data pipelines and flows
- Experience with developing software code and APIs in one or more languages such as Python and C#
- Experience designing and deploying large scale distributed data processing systems with one or more technologies such as MS SQL Server, PostgreSQL, MongoDB, Cassandra, Redis, Hadoop, Spark, Hive, Teradata, or MicroStrategy
- A high-level understanding of automation in a cloud environment (AWS experience preferred)
- Excellent communication, presentation, and problem-solving skills