Technology

Senior Data Scientist At Natural State

Natural State·Nairobi, kenya·Full Time·Remote
TechnologyFull TimeRemote
-

About the role

  • This Senior Data Scientist role is broad - we are looking for someone who can sit at the intersection of data engineering, ecological data science, and technical product ownership. As Senior Data Scientist, you will be responsible for day-to-day data management and the building of automated data processing pipelines and visualizations. You will also be involved in applying discriminative machine learning and ecological modeling approaches to processing and analyzing biodiversity data and designing and building ecological decision support tools within our Natural State Analytics platform.
  • This role owns data pipeline design, requirements, testing, and validation. Pipeline implementation may be carried out by you, the Technology team, or collaboratively depending on the project. This will include designing dynamic, structured field survey forms (in ODK) and data ingestion pipelines that ensure raw field data are cleaned, processed, and joined correctly to become analysis-ready datasets that can be exported for different end users. Your role will also include analyzing field datasets into decision-support metrics that can be displayed on platform dashboards and be used to generate project reports, under guidance from Project Delivery and Biometrics.
  • Our Project Delivery and Biometrics teams decide how data collection should be done, what quality checks matter, and what data dashboards need to show. You will turn these requirements into precise, buildable specifications for the Technology team to implement on the platform and then verify what gets built. You will not need to write all the backend code yourself, Natural State's Technology team owns most implementation, but you do need to be comfortable building data pipelines (in SQL) and writing and executing data analysis functions (in Python or similar). Most importantly, you need to understand the Natural State science database and principles deeply enough to specify exactly what should be built, ask the right questions before it's built, independently check the result once its built, and communicate that information clearly to external technical and non-technical audiences.
  • There is a lot of scope for growth within this role. We envision the first 6-12 months to be heavily focused on creating and improving data management pipelines. Once those systems become less time consuming to maintain, we hope you will bring your imagination and expertise to help us build data tools within the Natural State Analytics platform that inform better decision making and help us achieve our mission of Restoring the Natural World.

This role is for you if...

  • You have a strong technical background and are wanting to use your skills on real-world, applied biodiversity conservation and restoration projects.
  • You have ideas about how biodiversity data can be leveraged to make better decisions and you are excited to implement these.
  • You are detail-oriented and diligent, with a keen understanding of the importance of well-curated data and a passion for creating pipelines to support this.
  • You thrive in a remote work environment and are able to effectively manage your own workload without needing too much top-down direction.
  • You enjoy working in a small, high-performance team and pitching in where you're needed.

Requirements

  • Must have: 5+ years of work experience and a degree in data science, computer science, statistics, mathematics, quantitative ecology or a related field plus experience working with ecological/biodiversity data (e.g. camera trap images, passive acoustic monitoring recordings, vegetation surveys, animal surveys, species lists, soil carbon, biomass, remote sensing observation, climate etc.).
  • Must have: Strong Python for scientific data work - pandas or polars for data handling, plus the analysis stack (numpy, scipy, statsmodels, scikit-learn or equivalent). Comfortable writing validation scripts that catch schema mismatches early and explain clearly what broke.
  • Must have: Strong statistical reasoning, including an understanding of concepts such as confidence intervals, uncertainty estimation, GLMs, and discriminative machine learning models.
  • Must have: Confident in SQL, joins, CTEs, window functions.
  • Must have: Comfortable investigating data issues independently in pgAdmin, DBeaver, or similar.
  • Strongly desired: Able to handle spatial data in PostGIS, QGIS, and GeoPandas - raster algebra, coordinate reference systems and reprojection, vector versus raster, and the common ways location data breaks.
  • Strongly desired: Experience working with remote sensing data.
  • Nice to have: Experience with ODK, KoboToolbox, Survey123, or a comparable field data collection platform (ODK Central admin experience is a plus).

Key skills

BA/BSc/HND

At a glance

Company

Natural State

Location

Nairobi, kenya

Employment

Full Time

Experience

Valid until

Not specified

Created

September 11, 2026

More opportunities

Similar roles you might like

Senior Data Scientist Clean Cooking (payg Lpg) At Sun King (formerly Greenlight Planet)

Sun King (Formerly Greenlight Planet)

Nairobi, kenya

full-time

What Success Looks like Your work will be measured against the commercial metrics of the PAYG LPG business, not model metrics alone. In your first 12 months you will be expected to: Ship 2 - 3 data products into production use by commercial or operations teams, each with a measured impact on at least one of: activation rate, refill frequency, dormancy/churn, ARPU or cost-to-serve. Establish the BU's approach to pricing and promotion analysis elasticity estimates, uplift measurement and experiment design that commercial leaders actually use to make decisions. Build monitoring for the models you ship, with agreed retraining triggers and a clear owner for each. What you will be expected to do: Commercial Develop first-hand understanding of customer and commercial needs across our markets, including time in the field with sales agents and customers. Translate business problems into well-framed statistical questions, and present findings clearly to both technical and non-technical stakeholders up to BU leadership. Prioritise ruthlessly: identify where a data product will move activation, refill frequency, retention or unit economics, and be willing to say where it won't. Manage stakeholder expectations, product scope and delivery timelines for your own workstreams. Technical Design, build and evaluate machine learning models for business-critical use cases: churn/dormancy prediction, credit and payment-behaviour modelling, demand forecasting, anomaly detection and customer segmentation. Apply probabilistic and Bayesian methods to quantify uncertainty and support decisions under uncertainty e.g. pricing elasticity, promotion uplift and media/marketing effectiveness. Design and analyse experiments (A/B tests, geo tests, quasi-experiments) in field conditions where clean randomisation is often impossible. Perform rigorous exploratory analysis, feature engineering and data wrangling on large structured and semi-structured datasets. Partner with data and analytics engineering, who own pipelines and production infrastructure: you own the model from problem framing through validated, deployment-ready handoff, and jointly own monitoring once live. Track and communicate model performance; identify degradation and recommend retraining or redesign. Maintain clean, reproducible, well-documented code following team engineering standards. What this role is not about: Not a people-management role this is a senior IC position (a path to leading a small team may open as the function grows). Not an MLOps/platform role you'll work to production standards, but pipeline and deployment infrastructure is owned by MLOps Engineer. Not a reporting/BI role dashboarding exists in the analytics team; this role builds models and data products. You might be a strong candidate if you have: Degree in Computer Science, Statistics, Mathematics, Engineering, Economics or a closely related quantitative discipline. An advanced degree is a plus, not a requirement evidence of shipped impact matters more. Commercial A demonstrable track record of data products that measurably moved a business outcome you can walk us through the problem, the model, the decision it changed and the number it moved. Ability to listen to and empathise with customers and colleagues, and to identify the P&L impact of a proposed data product before building it. Strong communication and storytelling: you can carry a room of non-technical commercial leaders. Experience managing upwards product needs, trade-offs and timelines. Technical 5 - 8 years of hands-on experience in data science or applied ML roles, with at least 2 years owning data products end to end. Strong command of classical ML (gradient boosting, regression, clustering, ranking, time-series forecasting) and the judgment to know when simple beats sophisticated. Solid grounding in probabilistic modelling, Bayesian inference and uncertainty quantification, with working experience in a PPL such as PyMC or Stan. High proficiency in Python (the standard scientific stack) and strong SQL, including complex multi-table queries and window functions. Deep familiarity with model evaluation: cross-validation, calibration, and choosing business-aligned metrics over convenient ones. Experience with experiment design and statistical hypothesis testing. Comfortable working with cloud data warehouses and experiment-tracking tooling (we use AWS and MLflow; equivalents are fine). Strongly preferred Experience in PAYG, fintech lending, telco or other emerging-market consumer businesses you understand irregular incomes, mobile-money payment behaviour and thin, messy data. Nice to have Survival modelling, causal inference or marketing mix modelling (MMM). Operations research / optimisation exposure (routing, scheduling) relevant to our last-mile delivery problems. Familiarity with MLOps and model deployment on AWS (SageMaker, Lambda, ECS).

2 days ago

Full Stack Data Scientist Iii/iv At Idinsight

IDinsight

Nairobi, kenya

full-time

About the Role We are seeking candidates with a strong background in Python, deep applied expertise in one or more data science specialties (e.g., machine learning, LLMs/GenAI, optimization, geospatial analytics, MLOps, or full-stack engineering for data products), experience building and deploying solutions in production, and a passion for building solutions to difficult social problems. Most importantly, successful candidates should have the ability to learn and adapt quickly, and work independently to solve complex human and technological challenges. As a senior data scientist, you'll lead multiple projects as tech lead and be responsible for the performance of the solutions. Day-to-day work may include: Working with clients to understand their needs: Understanding their current processes and pain points, identifying which of these can be framed as tractable data science problems, and knowing when they can't, and even when it is, the solution must suit the task and resources available. Leading solution design and delivery end-to-end: Rolling up your sleeves as an individual contributor- writing production code, designing and evaluating methods and models, and building and deploying solutions (including APIs, UIs, CI/CD, etc.); while also working alongside other data scientists to shape the overall approach, synthesize findings, and communicate results. Establishing standards and best practices: Bringing experience in writing quality code, conducting code reviews, and providing feedback on technical and non-technical documentation. Setting the bar for engineering rigor and project management on the team. Coaching and mentoring: Upskilling junior data scientists on methods and approaches, and providing structured feedback on both technical and delivery skills. Providing thought leadership in your specialty: Leading the org through deep expertise in a subset of the following (machine learning, deep learning, GenAI/LLMs, NLP, optimization, geospatial analytics, backend development, frontend development, DevOps, or MLOps) and drawing on broad familiarity across methods to make sound judgment calls on which approach fits which problem. Shaping the sector: Identifying trends and gaps in the AI-for-Good space; proposing products or services that could address these gaps; and contributing thought leadership on how data science and AI can be deployed responsibly and for the right problem types. Moreover, professional development for our technical roles is essential for IDinsight's long-term impact. With support from IDinsight leadership, the employee will maintain self-directed professional development plans and will be given "stretch" opportunities designed to strengthen their professional skills. Real-time feedback and structured reviews are regularly provided to maximize each data scientist's expertise. IDinsight's entrepreneurial culture allows roles and career progression to be tailored to individual strengths, interests, and goals. Employees have the opportunity to increase responsibilities, and high performers will have the opportunity to move up in the organization along technical, managerial, or client-facing paths. Required Technical Qualifications Master's degree and 8 years of experience, or a PhD and 4 years of experience, as a data scientist working in Python. Demonstrated expertise in a subset of the following data science specialties: predictive modelling, machine learning, deep learning, GenAI/LLMs, NLP, optimization, or geospatial analytics. Intermediate-to-advanced Python skills / experience working on complex codebases. Working knowledge of at least one of: AI engineering, MLOps, backend development, DevOps, and frontend development Broad knowledge of advanced machine learning and data science methods. Strong foundations in statistics and probability. Proficiency in collaborative software development practices such as version control and code reviews Other required qualifications: Proven ability to work independently and with teams in a dynamic, multicultural environment. Experience leading technical teams to deliver complex data science or AI solutions, with a strong interest in mentoring, knowledge-sharing, presenting work and providing feedback to others. Strong oral and written communication skills in English. Professional proficiency in French (written and spoken) is a plus. Self-starter who will thrive while tackling new, unusual and unpredictable challenges. Deeply passionate about global development and improving lives in disadvantaged populations.

a month ago