Work with model researchers to define what “good data” means for our models, including quality metrics, validation checks, and acceptance thresholds
Explore open source datasets and create internal ones most suitable to build fundamental World Models
Build algorithms for automated data quality assessment, data domain mixtures, and domain adaptation from synthetic to real data.
Track datasets, metadata, provenance, and versions so experiments are reproducible and it’s clear what data went into which training and evaluation runs
Own CI/CD and development tooling for the data stack (GitHub, Python, PyTorch), and automate repetitive workflows to reduce friction
Track and optimize throughput, storage, and compute utilization across pipelines and related assets
PythonMachine LearningPyTorch+5 more
Showing 1 of 6 positions
About Reka
Reka, established in 1979, is a Danish company that designs, produces, and installs biomass boilers and complete combustion/heating plants. Their mission is to create sustainable energy solutions that are economical, environmentally friendly, and innovative. With over 40 years of experience, Reka is a leader in utilizing biomass resources like straw, woodchips, and sawdust for energy production, particularly in Denmark. They offer hand-fired boilers, automatic-fired boilers up to 4MW, and combustion plants up to 6500 kW, processing biomass with up to 55% moisture content. Reka is committed to improving existing boilers and expanding the types of fuel they can use.