soda-core
dictum
soda-core | dictum | |
---|---|---|
5 | 2 | |
1,765 | 19 | |
2.3% | - | |
8.9 | 0.0 | |
4 days ago | over 1 year ago | |
Python | Python | |
Apache License 2.0 | GNU General Public License v3.0 or later |
Stars - the number of stars that a project has on GitHub. Growth - month over month growth in stars.
Activity is a relative number indicating how actively a project is being developed. Recent commits have higher weight than older ones.
For example, an activity of 9.0 indicates that a project is amongst the top 10% of the most actively developed projects that we are tracking.
soda-core
- Looking for Unit Testing framework in Database Migration Process
-
Data profiling tools / approaches?
Tools like Soda Core could be really helpful for this. For example, it allows you to set up a change over time threshold which could take the form of: change avg last 3 for missing_count(column_name) < 20%
-
Data QC? Great Expectations?
You can give https://github.com/sodadata/soda-core - open source and (in my opinion) easy to get a lot of value with minimum effort.
- Show HN: Soda Core is now GA – Test data like you would test your code
-
Soda Core (OSS) is now GA! So, why should you add checks to your data pipelines?
Give Soda Core a try! It's really easy. If you only have 2 minutes, check out our docs or interactive demo (pretty cool no?). If you have a bit more time, install it and give it a spin! Want to look at it later? Star on Github. Got stuck? As in our Slack community.
dictum
-
Looking for feedback on my open-source project
TL;DR: Github repo | Documentation
-
Looking for feedback on my open-source Python library for Business Intelligence
I have been working on this project called Dictum. Here's the Github repo and Documentation. Looking for some feedback from BI people.
What are some alternatives?
great_expectations - Always know what to expect from your data.
sheetfu - Python library to interact with Google Sheets V4 API
dbt-data-reliability - dbt package that is part of Elementary, the dbt-native data observability solution for data & analytics engineers. Monitor your data pipelines in minutes. Available as self-hosted or cloud service with premium features.
metricflow - MetricFlow allows you to define, build, and maintain metrics in code.
cuallee - Possibly the fastest DataFrame-agnostic quality check library in town.
Apache Superset - Apache Superset is a Data Visualization and Data Exploration Platform [Moved to: https://github.com/apache/superset]
data-diff - Compare tables within or across databases
superset - Apache Superset is a Data Visualization and Data Exploration Platform
dbt-snowflake-monitoring - A dbt package from SELECT to help you monitor Snowflake performance and costs
Flight-Test-Data-Analytics-Module-01 - Code to support Module 01 of the Daedalus Aerospace Flight Test Data Analytics course.
pointblank - Data quality assessment and metadata reporting for data frames and database tables
gitbi - Lightweight BI app based on git repo