Data Scientist

Remote $81k–$138k middle 4 months ago full-time quality 8.2/10

Role in brief

Deel, a global payroll and HR platform, is seeking a Data Scientist to join its remote team. This role involves using data science and statistical methods to solve real-world problems, build and maintain datasets, and develop machine learning pipelines. Candidates with strong Python and SQL skills, experience in data/backend engineering, and a background in NLP and machine learning research implementation should consider applying.

PythonSQLdata sciencemachine learningNLPdata pipelining

About the role

This Data Scientist role at Deel focuses on applying data science and statistical techniques to solve practical problems. The work involves implementing functionalities that serve both internal and external customers, requiring the design, building, and maintenance of datasets. Key responsibilities also include data cleaning, modeling, feature engineering, and extraction to prepare data for analysis and algorithmic solutions.

A core part of this position is building end-to-end data and machine learning pipelines, ensuring that research is reproducible and can be deployed effectively. The successful candidate will collaborate closely with engineering, operations, and product management teams to deliver algorithmic solutions, integrating software engineering practices into research and infrastructure.

Success in this role means producing high-quality, clean, and maintainable research and code. It requires an independent thinker who can work autonomously to solve problems, while also effectively communicating complex ideas to diverse audiences. The role emphasizes a hands-on approach to applying statistical methods and developing AI solutions across various domains.

The salary for this position ranges from $80,500 to $138,000 USD annually.

Skills that matter here

  • Python: This role requires high proficiency in Python for data science tasks, implementing research, and developing AI solutions.
  • SQL: Hands-on experience with SQL is essential for designing, building, and maintaining datasets.
  • data science: The position involves applying data science techniques to solve real-world problems and develop insights from large datasets.
  • machine learning: Experience in implementing machine learning algorithms and building end-to-end machine learning pipelines is a core requirement.
  • NLP: A background in text processing and Natural Language Processing (NLP) is necessary for implementing research and algorithms.
  • data pipelining: Strong skills in data pipelining are crucial for building and maintaining the infrastructure for data and machine learning processes.

Who this role suits

  • Someone with a background in data or backend engineering, ideally with experience in production environments.
  • An individual who thrives on implementing research and algorithms in Python, particularly in areas like information retrieval and NLP.
  • A person who can communicate complex technical concepts clearly and collaborate effectively with diverse teams.
  • An independent thinker who is capable of working autonomously to solve problems and apply statistical methods to data.

From the employer

  • Solve real world problems using Data Science and statistical techniques
  • Implement functionality which can be served in production for internal customers as well as external customers
  • Designing, building and maintaining data sets
  • Data cleaning & modelling
  • Feature engineering
  • Feature extraction
  • Building end-to-end data & machine learning pipelines
  • Conduct reproducible research
  • Collaborate with Engineering, Operations, Product Management and other functions in the company to deliver algorithmic solutions
  • Apply software engineering practices in our code that implements our research and its infrastructure
  • Produce high quality, clean, maintainable reproducible research and code
  • High proficiency in Python and its data science stack.
  • Background and experience in data/backend engineering, ideally in production environments (3+ years) (Mid/Senior Level role)
  • Background and hands-on experience (2+ years) in implementing research and algorithms in Python, specifically in information retrieval, text processing, NLP, and machine learning
  • Experience with developing AI solutions across a variety of domains
  • Track record of good written and verbal communication of complex things in a simple way as well as ability to collaborate well with people from different backgrounds and professions
  • A Bachelor’s degree or higher
  • Must have hands on experience working with SQL
  • Must have hands on experience working with Python (Preferably with Pandas)
  • Must be strong at applying statistical methods to data
  • Must be strong with data pipelining
  • Must be an independent thinker and have the ability to work independently to solve problems
  • Stock grant opportunities dependent on your role, employment status and location
  • Additional perks and benefits based on your employment status and country
  • The flexibility of remote work, including optional WeWork access

Questions about this role

What is the remote work policy for this role?

This is a fully remote position, offering the flexibility of remote work and optional WeWork access.

What level of seniority is expected for this position?

This is a middle-senior level role, requiring 3+ years of experience in data/backend engineering and 2+ years in implementing research and algorithms in Python.

What are the core technical skills required for this role?

Candidates must have high proficiency in Python and its data science stack, hands-on experience with SQL, and strong skills in applying statistical methods and data pipelining.

Similar jobs

Before you apply

  • Legitimate employers never ask you to pay anything to apply or get hired.
  • Never share seed phrases or private keys. No real job needs them.
  • Do not install software ("test tasks", "trading tools", "video call clients") sent during hiring.
  • Check that the application page's domain really belongs to Deel.