Skip to content
View margitaii's full-sized avatar
  • Vienna, AT

Block or report margitaii

Block user

Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.

You must be logged in to block users.

Maximum 250 characters. Please don’t include any personal information such as legal names or email addresses. Markdown is supported. This note will only be visible to you.
Report abuse

Contact GitHub support about this user’s behavior. Learn more about reporting abuse.

Report abuse
Showing results

An orchestration platform for the development, production, and observation of data assets.

Python 16,032 2,249 Updated Aug 19, 2026

A dynamic data completeness and accuracy library at enterprise scale for Apache Spark

Scala 30 9 Updated May 13, 2026

A PyTorch implementation of Convolutional Sequence Embedding Recommendation Model (Caser)

Python 276 69 Updated Feb 4, 2021

A Matlab implementation of Convolutional Sequence Embedding Recommendation Model (Caser)

MATLAB 77 23 Updated Jul 15, 2018

Python API for Deequ

Jupyter Notebook 826 157 Updated Jul 21, 2026

Jenga is an experimentation library that allows data science practititioners and researchers to study the effect of common data corruptions (e.g., missing values, broken character encodings) on the…

Jupyter Notebook 43 7 Updated Jun 21, 2023

Smart Automation Tool for building modern Data Lakes and Data Pipelines

Scala 130 26 Updated Aug 20, 2026

Automated data quality suggestions and analysis with Deequ on AWS Glue

Scala 93 23 Updated Dec 29, 2022

R package implementing the SA-CCR based on the CRR2 Regulation

R 7 14 Updated Jul 5, 2021

Deequ is a library built on top of Apache Spark for defining "unit tests for data", which measure data quality in large datasets.

Scala 3,641 584 Updated Jul 21, 2026