Careers / Data & AI
Data Engineer
Full-timeRemote (SA)Data & AI
You will build the pipelines and warehouses behind our data and analytics engagements: ingesting messy operational data, modelling it properly, and serving it to dashboards executives actually use. Most projects run on cloud-managed warehouses and orchestrated Python.
What you will do
- Design and build ELT pipelines from client systems (ERPs, operational databases, third-party APIs) into cloud warehouses
- Model warehouse schemas with clear staging, core, and mart layers
- Write and maintain data quality tests so bad loads fail loudly instead of silently
- Set up orchestration, alerting, and cost monitoring for every pipeline you ship
- Work with analysts and clients to translate reporting questions into models
- Document sources, transformations, and definitions so clients can self-serve after handover
What you bring
- 4+ years in data engineering with strong SQL and production Python
- Hands-on experience with at least one cloud data warehouse and one orchestration tool
- Solid dimensional modelling skills; you can defend a fact/dimension design under questioning
- Experience ingesting from imperfect sources: CSV drops, legacy databases, rate-limited APIs
- Comfort owning pipelines in production, including on-call for the ones you build
- POPIA awareness: you know what personal data is and how to handle it in a pipeline
Nice to have
- dbt experience with tests and documentation in the repo
- Streaming or CDC experience for near-real-time loads
- Exposure to ML feature pipelines or model deployment
- Experience presenting findings directly to client stakeholders
The process
Portfolio review, one technical conversation about work you have shipped, and a session with the pod you would join. About two weeks end to end, and every application gets a human reply.
