Responsibilities
- Develop and operate data pipelines
- Build data layers in OneDataLake (ODL) and Central Data Store Consumer (CDSC)
- Balance business and IT demands to build modern and efficient pipelines in public cloud environments
- Integrate data preparations into managed processes
- Implement data integrations in Data Vault
- Implement data exports and data operations
Requirements
- Experience with Python, PySpark, SQL
- Familiarity with data pipeline development and operations
- Ability to work with Gitlab (Magenta CI/CD)
- Experience implementing data integrations in Data Vault
- Ability to build data layers in a central data lake environment
Nice to Have
- Experience with libraries such as Pandas, Matplotlib, Seaborn, Bokeh, Plotly
- Background in cloud-based data engineering
- Involvement in generative AI use cases such as NPS-X check or Telekom-specific large language model (LLM) adoptions
Additional Information
- The role involves enabling and supporting generative AI use cases including NPS analytics and dashboards, NPS Feedback analytics, and adhoc analytics for CX Outerloops
- The team establishes a single source of truth around the customer and ensures unified data flow of TV-relevant data across systems
- Focus on data analytics use cases for digital touchpoints of the German NatCo
- The role contributes to establishing a common way of using generative AI in the organization