Data Science | AI Enthusiast | Apache Spark Developer
I am a Data Scientist with strong experience in data engineering, especially Apache Spark, PySpark, and Azure Databricks. With a Master’s degree from the Technical University of Dortmund and over 3 years of experience, I’ve worked on scalable data pipelines, spark optimization, anomaly detection, time-series modeling. I’m passionate about designing robust, high‑performance data systems and applying AI through RAG architectures, agents, and MCP to deliver intelligent, production‑ready solutions.
- Current Role: Data Scientist/Engineer at Daimler AG
- Past Experience: Data Engineer Intern at Henkel, Machine Learning and Data Engineering roles in industry and academia, including Intel and research institutes
- Focus Areas: Feature Engineering, Time series modelling, Anomaly detection, ETL workflows, Generative AI, LLM-based applications (LangChain, RAG)
- Languages: PySpark, Python, R, SQL
- ML/AI Tools: Scikit-learn, PyTorch, LSTM/GRU, Autoencoders, Gurobi
- Big Data & Cloud: Azure Databricks, Apache Spark, GCP, Snowflake, BigQuery
- Visualization: Power BI, Tableau, Plotly, Looker Studio, Matplotlib, Seaborn
- Databases: Azure Data Lake, MySQL, PostgreSQL, MongoDB