versatile-data-kit
terraform-cdk
Our great sponsors
versatile-data-kit | terraform-cdk | |
---|---|---|
52 | 104 | |
410 | 4,727 | |
2.4% | 1.2% | |
9.7 | 9.8 | |
6 days ago | about 24 hours ago | |
Python | TypeScript | |
Apache License 2.0 | Mozilla Public License 2.0 |
Stars - the number of stars that a project has on GitHub. Growth - month over month growth in stars.
Activity is a relative number indicating how actively a project is being developed. Recent commits have higher weight than older ones.
For example, an activity of 9.0 indicates that a project is amongst the top 10% of the most actively developed projects that we are tracking.
versatile-data-kit
-
Looking for a data blogger
Here's the project: https://github.com/vmware/versatile-data-kit
-
Need advice on ETL tool
I don't really know if this would work for you because the UI is not functional yet, but a very simple REST API ingestion example here, there's one for csv too https://github.com/vmware/versatile-data-kit/wiki/Ingesting-data-from-REST-API-into-Database I can't imagine a simpler way unless it's really drag and drop.
-
If dbt is the "T" part of an "ELT", what do you use for "EL"?
I work at VMware and we use one tool for the whole ELT, it was made internally as there was no good alternative at the time and now we opensourced it, here it is: https://github.com/vmware/versatile-data-kit
-
Best way to fix errors in my data?
With my team we created csv ingestion plugin described here, maybe you want to try it out: https://github.com/vmware/versatile-data-kit/wiki/Ingesting-local-CSV-file-into-Database
-
What Orchestration Tool do you use for batch ETL/ELT?
We use Versatile Data Kit for batch data job orchestration (https://github.com/vmware/versatile-data-kit)
-
Dear, pipeline builders! Which step in your role is the most time consuming?
"suggestions on how to reduce the time spent on initially generating and adjusting the code" is using some tools that automate ELT. Here's one open-source tool I'm working on with my team: https://github.com/vmware/versatile-data-kit
-
Problem definition / vibe check for a repo
here's the repo: https://github.com/vmware/versatile-data-kit
-
Can we take a moment to appreciate how much of dataengineering is open source?
If you wish to contribute, projects usually have good first issues: https://github.com/vmware/versatile-data-kit/labels/good%20first%20issue If you wish to learn, check out examples: https://github.com/vmware/versatile-data-kit/tree/main/examples
-
ETL question (noob)
Have you heard about versatile data kit (https://github.com/vmware/versatile-data-kit)? I think it meets your needs perfectly:
-
DE Open Source
Versatile Data Kit is a framework to bBuild, run and manage your data pipelines with Python or SQL on any cloud https://github.com/vmware/versatile-data-kit here's a list of good first issues: https://github.com/vmware/versatile-data-kit/issues?q=is%3Aissue+is%3Aopen+label%3A%22good+first+issue%22 Join our slack channel to connect with our team: https://cloud-native.slack.com/archives/C033PSLKCPR
terraform-cdk
-
Learning Go by examples: part 12 - Deploy Go apps in Go with CDK for Terraform (CDKTF)
At first I tested it to deploy an OVHcloud Managed Kubernetes Service (MKS) with a Node Pool. And step by step, it worked. I even created a Pull Request (PR) in the terraform-cdk repository to add it as an example βΊοΈ.
-
AWS CDK For Noobs: Deploying NextJS Apps
I'll be trying more sample app deployments with CDK and maybe even explore CDK for Terraform.
-
Show HN: Winglang β a new Cloud-Oriented programming language
You can use CDK with other providers using https://github.com/hashicorp/terraform-cdk
In my experience, CDK is far better than Pulumi, especially if you're mostly going to be using AWS.
- Terraform CDK
-
Why is Kubernetes adoption so hard?
I, personally, prefer Crossplane Composite Functions on top of CDK8S, but had dropped CDKTF due to bloat. You can actually manage Kubernetes updates/upgrade lifecycle with Crossplane, as well.
- Cloud, Why So Difficult?
- What are some harsh truths that r/devops needs to hear?
-
Backend engineers that don't like JavaScript
I was going to recommend Pulumi, but looks like CDK for Terraform is still being kept up to date.
-
Should i migrate from Kustomize to Helm?
Avoid Pulumi, get directly to source and use https://github.com/hashicorp/terraform-cdk
- AWS IAM Roles, a tale of unnecessary complexity
What are some alternatives?
data-engineering-zoomcamp - Free Data Engineering course!
Pulumi - Pulumi - Infrastructure as Code in any programming language. Build infrastructure intuitively on any cloud using familiar languages π
Mage - π§ The modern replacement for Airflow. Mage is an open-source data pipeline tool for transforming and integrating data. https://github.com/mage-ai/mage-ai
terragrunt - Terragrunt is a thin wrapper for Terraform that provides extra tools for working with multiple Terraform modules.
quadratic - Quadratic | Data Science Spreadsheet with Python & SQL
crossplane - The Cloud Native Control Plane
pyramid-jsonapi - Auto-build JSON API from sqlalchemy models using the pyramid framework
copilot-cli - The AWS Copilot CLI is a tool for developers to build, release and operate production ready containerized applications on AWS App Runner or Amazon ECS on AWS Fargate.
dbt-data-reliability - dbt package that is part of Elementary, the dbt-native data observability solution for data & analytics engineers. Monitor your data pipelines in minutes. Available as self-hosted or cloud service with premium features.
cdk8s - Define Kubernetes native apps and abstractions using object-oriented programming
hamilton - A scalable general purpose micro-framework for defining dataflows. THIS REPOSITORY HAS BEEN MOVED TO www.github.com/dagworks-inc/hamilton
aws-cdk-local - Thin wrapper script for using the AWS CDK CLI with LocalStack