Kpler Logo

Kpler

Data Scientist

Posted 18 Days Ago
Remote or Hybrid
Hiring Remotely in Greece
Junior
Remote or Hybrid
Hiring Remotely in Greece
Junior
Design, implement, and evaluate production ML models for ETA, destination, and product estimation using maritime AIS and geospatial features. Build dbt features orchestrated by Airflow, run structured experiments with the kpler-ml framework, log and compare runs in MLflow, monitor model drift, and collaborate with market analysts to prioritize improvements.
The summary above was generated by AI
At Kpler, we are dedicated to helping our clients navigate complex markets with ease. By simplifying global trade information and providing valuable insights, we empower organisations to make informed decisions in commodities, energy, and maritime sectors.
 
Since our founding in 2014, we have focused on delivering top-tier intelligence through user-friendly platforms. Our team of over 850 experts from 69 countries works tirelessly to transform intricate data into actionable strategies, ensuring our clients stay ahead in a dynamic market landscape. Join us to leverage cutting-edge innovation for impactful results and experience unparalleled support on your journey to success.
 

Data Scientist

The Commodities tribe at Kpler runs production ML models that predict what cargo a vessel is carrying (Product Estimation) and where in-transit vessels are headed (Destination Forecast), and where they are expected to arrive (ETA)— across LNG, DRY, LPG, and LIQUIDS. These predictions feed directly into Kpler's cargo intelligence platform, consumed by market analysts, trading desks, and external customers worldwide.

You will own the science behind these models: designing and evaluating features from maritime AIS data, H3 geospatial routing distributions, transit statistics, and commodity-specific signals; running structured experiments on ML Flow based platform; and pushing the accuracy, coverage, and reliability of predictions forward.

You are not handed a Jupyter notebook and a dataset. You work in a production system with real-time inference running every 1–3 hours across 4 commodity types, and your model changes need to be validated against a running parallel baseline before they go live. The new platform is being built specifically to make the experiment loop fast enough that this level of rigour does not slow you down.

Key Responsibilities

  • Own the feature engineering roadmap for ETA & Destination Forecast across all 4 commodity types — propose and implement new features as dbt models using Airflow to orchestrate the data pipelines, and validate their impact through structured experiments.
  • Design and run experiments using kpler-ml framework, logging all runs from train to evaluation to MLflow and producing structured comparison reports against the production baseline before any promotion.
  • Work directly with Commodities Market Analysts and product stakeholders to understand where prediction quality matters most commercially — and use that to prioritise the experiment backlog.
  • Contribute to the drift monitoring setup — validate PSI/KS thresholds using MLFlow against historical inference batches; define what constitutes a meaningful drift signal for PE and DF specifically.
  • Document experiment decisions in MLflow and Confluence documents — the experiment history is a first-class artifact, not an afterthought.

Experience & Background

  •  2+ years applying ML to real-world production problems — not research or hackathon work, but models running in production with real consequences for errors 

  •  Experience with geospatial or sequential data — vessel trajectories, routing patterns, H3/S2 grid systems, or equivalent spatial representations

  •  Python proficiency at a level sufficient to implement new features, write dbt models, and script experiments — not just use notebooks

  • Familiarity with MLflow or equivalent experiment tracking (Weights & Biases, Neptune, etc.)

    Desirable:

  • Domain knowledge of maritime shipping, commodity trading, or cargo intelligence — understanding what a port call sequence or a vessel's draught profile means physically, not just statistically

  • Familiarity with Redshift or columnar warehouses for large-scale feature queries and dbt (authoring or reading SQL models)


We are a dynamic company dedicated to nurturing connections and innovating solutions to tackle market challenges head-on. If you thrive on customer satisfaction and turning ideas into reality, then you’ve found your ideal destination. Are you ready to embark on this exciting journey with us?
 
We make things happen
We act decisively and with purpose, going the extra mile.
 
We build
together
We foster relationships and develop creative solutions to address market challenges.
 
We are here to help
We are accessible and supportive to colleagues and clients with a friendly approach.
 
 
Our People Pledge
 
Don’t meet every single requirement? Research shows that women and people of color are less likely than others to apply if they feel like they don’t match 100% of the job requirements. Don’t let the confidence gap stand in your way, we’d love to hear from you! We understand that experience comes in many different forms and are dedicated to adding new perspectives to the team.
 
Kpler is committed to providing a fair, inclusive and diverse work-environment. We believe that different perspectives lead to better ideas, and better ideas allow us to better understand the needs and interests of our diverse, global community. We welcome people of different backgrounds, experiences, abilities and perspectives and are an equal opportunity employer.
 
 
 
By applying, I confirm that I have read and accept the Staff Privacy Notice

Similar Jobs

9 Days Ago
Remote
Mid level
Mid level
Mobile • Software
Develop, improve, and deploy machine learning models for large-scale ad-tech systems; explore and integrate new data sources; research ML solutions; collaborate with engineering to productionize algorithms; and design and evaluate A/B tests and monitoring dashboards.
Top Skills: AirflowCatboostDockerHadoopLarge-Scale OptimizationLightgbmNeural NetworksPythonReinforcement LearningSparkSQL
3 Days Ago
Remote
Senior level
Senior level
Information Technology • Internet of Things • Financial Services
Lead data science projects end-to-end: define high-impact problems, build and productionize models (recommendations, uplift, forecasting), own experimentation and measurement, operate reliable pipelines, and mentor the team to drive business metric improvements.
Top Skills: AirbyteAWSBigQueryClaudeDbtDockerGCPGitGithub ActionsHexPostgresPythonSnowflakeSQLTerraformTypescript
3 Days Ago
Remote
Canada
Senior level
Senior level
Fintech • Mobile • Payments • Software
Proactively explore RevenueCat data to identify customer problems and opportunities. Translate ambiguous product questions into analyses and production-grade predictive/descriptive models. Partner with Product, Engineering, and Analytics to shape roadmaps, define experimentation and benchmarking approaches, deploy and iterate on models powering customer-facing features, and communicate insights across technical and non-technical audiences.
Top Skills: AWSDbtPostgresPythonSnowflakeSQL

What you need to know about the Calgary Tech Scene

Employees can spend up to one-third of their life at work, so choosing the right company is crucial, not just for the job itself but for the company culture as well. While startups often offer dynamic culture and growth opportunities, large corporations provide benefits like career development and networking, especially appealing to recent graduates. Fortunately, Calgary stands out as a hub for both, recognized as one of Startup Genome's Top 100 Emerging Ecosystems, while also playing host to a number of multinational enterprises. In Calgary, job seekers can find a wide range of opportunities.

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account