zingg
yt-channels-DS-AI-ML-CS
Our great sponsors
zingg | yt-channels-DS-AI-ML-CS | |
---|---|---|
23 | 5 | |
877 | 1,254 | |
2.3% | - | |
9.3 | 0.0 | |
7 days ago | over 1 year ago | |
Java | ||
GNU Affero General Public License v3.0 | - |
Stars - the number of stars that a project has on GitHub. Growth - month over month growth in stars.
Activity is a relative number indicating how actively a project is being developed. Recent commits have higher weight than older ones.
For example, an activity of 9.0 indicates that a project is amongst the top 10% of the most actively developed projects that we are tracking.
zingg
-
Ask HN: What is the most impactful thing you've ever built?
As part of my data consulting, I struggled with identity resolution and started working on scalable no code identity resolution - https://github.com/zinggAI/zingg/ . It has pushed my limits as a software engineer and product builder, and I had to do a lot of learning to build it. Its cool to see people use Zingg in their workflows and save months of working on custom solutions. Big highlight has been North Carolina Open Campaign Data https://crossroads-cx.medium.com/building-open-access-to-nc-...
-
How to find open source data science python projects to contribute to?
Check https://github.com/zinggAI/zingg/. We recently added Python to our stack and are looking for help with building dbt-zingg python models, databricks-zingg python notebooks, python api, building a python based front end etc.
- Merging datasets
-
is it possible to "fuzzy match" or dedupe columns in Redshift?
If you are open to using a framework for this, check Zingg at https://github.com/zinggAI/zingg. It connects to Redshift, snowflake and other warehouses and can handle multiple columns
-
Show HN: Zingg – open-source entity resolution for single source of truth
Thanks for your support. Yes we do ship with some examples and their models which can be run out of the box. We have 3 customer demographic datasets and an ecommerce items matching across Google and Amazon. You can check them here https://github.com/zinggAI/zingg/tree/main/examples
-
Question about Github Referring Sites
I have an open source project hosted at https://github.com/zinggAI/zingg/.
- How do I promote the project appropriately?
-
GitHub Java Projects to Contribute
Check Zingg out at https://github.com/zinggAI/zingg and let me know if you would like to contribute
-
Match over 1 GB of data with inconsistent names
This is interesting, would love to get your feedback on Zingg(https://github.com/zinggAI/zingg) if you are upto it. Thanks!
-
Open source entity resolution - need your feedback!
I have released an open source entity resolution tool Zingg(https://github.com/zinggAI/zingg). Zingg uses Spark and ML to build single source of truth directly in the warehouse or the datalake. Would love to hear from the Reddit folks here what they think about it - do you find it useful? what can I do to make it better? any advice on the problem or the solution?
yt-channels-DS-AI-ML-CS
- List of computer science learning sites
-
Any Really Good Computer Science or Coding Channels on YT?
Check out this list I compiled a while back. https://github.com/benthecoder/yt-channels-DS-AI-ML-CS. Hope it helps
- Which blogs/youtube channels should I follow as a beginner in Data sciences?
-
I compiled a list of YouTube channels for Data Science, Python, Software engineering, Machine Learning and more
Github repo: https://github.com/benthecoder/yt-channels-DS-AI-ML-CS Github pages: https://benthecoder.github.io/yt-channels-DS-AI-ML-CS/
- A List of YouTube Channels for Data Science, ML, AI, Programming, and More
What are some alternatives?
splink - Fast, accurate and scalable probabilistic data linkage with support for multiple SQL backends
awesome-chatgpt - 🤖 Awesome list for ChatGPT — an artificial intelligence chatbot developed by OpenAI
clrs
Data-Science-Machine-Learning-Project-with-Source-Code - Data Science and Machine Learning projects with source code.
rumble - ⛈️ RumbleDB 1.21.0 "Hawthorn blossom" 🌳 for Apache Spark | Run queries on your large-scale, messy JSON-like data (JSON, text, CSV, Parquet, ROOT, AVRO, SVM...) | No install required (just a jar to download) | Declarative Machine Learning and more
Computer-Science-Resources - A list of resources in different fields of Computer Science
CLRS - Algorithms implementation in C++ and solutions of questions (both code and math proof) from “Introduction to Algorithms” (3e) (CLRS) in LaTeX.
Mage - 🧙 The modern replacement for Airflow. Mage is an open-source data pipeline tool for transforming and integrating data. https://github.com/mage-ai/mage-ai
skipledger - Differential privacy solution for maintaining and exposing information from evolving, append-only journals / ledgers.
awesome-bigdata - A curated list of awesome big data frameworks, ressources and other awesomeness.
automount - Simple devd(8) based automounter for FreeBSD
awesome-javascript-learning - A tiny list limited to the best JavaScript Learning Resources