Seek logo
Seek
Posted 11 days agoVerified live 2d ago

Data Scientist - US Remote

Brief overview

Remote
UndergradOr in progress
$105k–$120k/yrStated range
1+ yrsMinimum
PythonRegressionClassificationClusteringData Preparation and Feature EngineeringSQLData Visualization and DashboardingPySparkSnowParkGitGitLab CISnowflakeDatabricksAmazon RedshiftMachine Learning Performance Optimization

About the company

Seek's Insight Cloud is the Analytics App Store for your Business.

Job description

Summary

Seek is redefining how businesses harness analytics through an ML cloud platform and app store. The Data Scientist will analyze complex datasets, develop and optimize analytical and machine learning models, build data workflows, and collaborate with BI developers, engineers, scientists, and stakeholders to deliver actionable insights through Seek’s Insight Cloud.

Responsibilities

  • Analyze large, sparse, and messy datasets to extract meaningful patterns, trends, and insights using advanced statistical methods and machine learning algorithms
  • Construct and maintain data processing workflows, ensuring efficient mapping and transformation of source data
  • Perform data enrichment by sourcing additional datasets and integrating them with publisher data to enhance analytical depth and accuracy
  • Develop complex queries and scripts to transform and model data, tailoring configurations and fine-tuning parameters for optimal performance
  • Translate business challenges into analytical frameworks, proposing and implementing data-driven solutions that have a tangible impact on our subscribers’ decision-making processes
  • Assist our BI Developers in creating interactive dashboards that provide intuitive visualizations of data insights, enabling users to derive clear, actionable information
  • Work closely with BI developers, data engineers, other data scientists and business stakeholders to ensure that models are aligned with business goals and technical requirements
  • Write clear and comprehensive documentation for models, pipelines, and workflows to facilitate collaboration and reproducibility

Skills

  • 1+ year of experience in data science or data analytics
  • Proficiency in common data science Python libraries for data transformation, regression, classification, and clustering
  • Experience in data prep and feature engineering including cleaning, normalizing, transforming, and selecting relevant features from raw data
  • Proficient SQL skills: joins, subqueries, window functions, common table expressions (CTEs), aggregations, complex string manipulations, stored procedures and functions
  • Experience with data visualization tools and techniques or dashboarding products
  • Excellent problem-solving and creative thinking skills with ability to drive complex data analysis and development forward independently
  • Strong communication skills with the ability to present data and articulate complex technical concepts clearly and concisely
  • Bachelor's degree in data science, statistics, computer science, or a related field
  • This is an entirely remote position for candidates based in the US
  • We do not offer visa sponsorship at this time
  • Experience with PySpark or SnowPark for Python, SQL user-defined functions and table functions (UDFs & UDTFs)
  • Familiar with version control like GitHub or Gitlab and production deployment pipelines
  • Experience working with large-scale datasets directly within cloud data warehouses such as Snowflake, DataBricks, or Redshift
  • ML performance optimization experience: efficient computation, parallel processing, hyperparameter turning, model pruning, and quantization

Qualifications

Must Haves

  • 1+ year of experience in data science or data analytics
  • Proficiency in common data science Python libraries for data transformation, regression, classification, and clustering
  • Experience in data prep and feature engineering including cleaning, normalizing, transforming, and selecting relevant features from raw data
  • Proficient SQL skills: joins, subqueries, window functions, common table expressions (CTEs), aggregations, complex string manipulations, stored procedures and functions
  • Experience with data visualization tools and techniques or dashboarding products
  • Excellent problem-solving and creative thinking skills with ability to drive complex data analysis and development forward independently
  • Strong communication skills with the ability to present data and articulate complex technical concepts clearly and concisely
  • Bachelor's degree in data science, statistics, computer science, or a related field
  • This is an entirely remote position for candidates based in the US
  • We do not offer visa sponsorship at this time

Nice to Haves

  • Experience with PySpark or SnowPark for Python, SQL user-defined functions and table functions (UDFs & UDTFs)
  • Familiar with version control like GitHub or Gitlab and production deployment pipelines
  • Experience working with large-scale datasets directly within cloud data warehouses such as Snowflake, DataBricks, or Redshift
  • ML performance optimization experience: efficient computation, parallel processing, hyperparameter turning, model pruning, and quantization

Benefits

  • Medical Insurance
  • Dental Insurance
  • PTO
  • 16 annual company holidays
  • 401K
  • Vision Insurance
  • This is an entirely remote position for candidates based in the US.

More jobs like this