gnomad-browser
grouparoo
Our great sponsors
gnomad-browser | grouparoo | |
---|---|---|
15 | 27 | |
78 | 607 | |
- | - | |
9.7 | 9.9 | |
3 days ago | about 2 years ago | |
TypeScript | JavaScript | |
MIT License | MIT License |
Stars - the number of stars that a project has on GitHub. Growth - month over month growth in stars.
Activity is a relative number indicating how actively a project is being developed. Recent commits have higher weight than older ones.
For example, an activity of 9.0 indicates that a project is amongst the top 10% of the most actively developed projects that we are tracking.
gnomad-browser
- All identified polymorphisms in a given gene, how to find?
- AskScience AMA Series: We're human genetics researchers here to discuss connections between people in different geographical regions. Ask us anything!
-
Converting 23&Me raw data into a format usable by Admixtools 2
try one of these: GnomAD (https://gnomad.broadinstitute.org/) 1000 Genomes (http://browser.1000genomes.org) dbSNP (http://www.ncbi.nlm.nih.gov/snp)
-
What is the maximum number of human?
Maybe you can ask the opposite question; what are the bounds of of a functional human being. https://gnomad.broadinstitute.org/ GnomAD is a aggregation of healthy human genetic sequences which was primarily built on the aggregated control groups of many genetic sequencing studies. There are studies of this data analysing the co-occurrence of variants in gnomAD which may help.
-
Insights from personal sequencing data I can explore.
Maybe something like this? https://promethease.com/ Clinvar for variants that might be of clinical relevance. https://gnomad.broadinstitute.org/ for allele frequencies & some info about variants.
-
What are some non-pathogenic alleles of the SNCA gene, or how do I find them?
You could look at aggregation databases such as gnomad https://gnomad.broadinstitute.org/ anything with a frequency incompatible with the disease is likely non pathogenic
-
Ask HN: Who is hiring? (March 2022)
Broad Institute of MIT and Harvard | Cambridge, MA | Frontend Software Engineer | REMOTE or HYBRID (New England area)
We are hiring a frontend developer to help lead the next phase of the gnomAD browser, a web application for displaying the world's largest collection of human genome/exome sequences. https://gnomad.broadinstitute.org. Looking for applicants who are excited about data visualization and designing complex interfaces for scientific research.
Apply here: http://broad.io/cq7dw8
-
Ask HN: Who is hiring? (February 2022)
Broad Institute of MIT and Harvard | New England | Software Engineer | REMOTE/HYBRID
Our team is focused on building the tools necessary to visualize and interpret massive data sets of human genetic variation and functional genomic information. We have developed gnomAD (https://gnomad.broadinstitute.org), the world’s largest public reference dataset of human exomes and genomes. gnomAD has become one of the most widely used resources in the field, and is now the default reference database for virtually all clinical interpretation pipelines, as well as a standard analysis resource for a wide variety of genetic and biological studies. We estimate gnomAD has contributed to the clinical diagnosis of over 2 million patients with genetic disorders.
Your role will be to maintain the gnomAD browser, our open source web application for exploring gnomAD and related datasets, and develop new scientific functionality as we continue to grow to over 1 million human samples. You will work with a team of software engineers, computational biologists and clinical and research users to develop new features and visualizations that incorporate user feedback. Software engineering skills and an interest in user interface design and data visualization are key. Basic familiarity with genomics and DNA sequencing data is preferred, but not required. Most importantly, the ideal candidate will have enthusiasm for playing a critical role in a team-oriented project and learning new domains.
Minimum Requirements
- Ask HN: How to be my own genetic disease researcher for my partner?
-
How to check if a discovered mutation is novel or was discovered before ?
If you're talking about humans, start with gnomAD: https://gnomad.broadinstitute.org/
grouparoo
-
Reverse ETL recommendations?
Reverse ETL is on AirByte's roadmap under the "Future / Not prioritized" section. I wanted to use Grouparoo as a short term solution, but the repo was archived and I think they stopped taking new cloud customers (unknown if this is wrong/outdated).
-
Reference Data Stack for Data-Driven Startups
There are other tools that we will have to adopt in the future but haven’t yet due to lack of necessity. Specifically, one category that is popular in modern data stacks is Reverse ETL (Hightouch, Census, or Grouparoo). We currently don’t have a usecase for piping data back into 3rd party tools but it will definitely come up in the future.
-
Data pipeline suggestions
Reverse ETL: Grouparoo, Castled
-
Is Reverse ETL a new product or a new ETL/ELT feature?
Grouparoo, the open source Reverse ETL tool we are building, does all of these things. https://www.grouparoo.com
-
Where can I find free data engineering ( big data) projects online?
Ingestion / ETL: Airbyte, Singer, Jitsu Transformation: dbt Orchestration: Airflow, Dagster Testing: GreatExpectations Observability: Monosi Reverse ETL: Grouparoo, Castled Visualization: Lightdash, Superset
- Invite your company
-
Ask HN: Who is hiring? (December 2021)
Grouparoo | Remote (US) | Remote-OK | https://www.grouparoo.com
Grouparoo is a venture-backed software company building open source data tools that make data reliable, accessible, and actionable. We’re empowering teams to make great customer experiences, driven by data. While engineering teams have gotten good at storing and generating data about their customers, it’s rare that this data is used to its full potential in external applications. Grouparoo makes these integrations easy by providing a framework for defining your customer data and reliably syncing it to external tools.
To learn more about who we are, our engineering culture, and whether this is the right place for you, read our Key Values profile: https://www.keyvalues.com/grouparoo
Here are our open roles:
- Senior Backend / Lead Engineer: https://jobs.lever.co/grouparoo/6ba485d1-a5a4-41f0-9fa5-920a...
- Developer Advocate: https://jobs.lever.co/grouparoo/5e1531b4-7ec8-4c10-8e52-fc23...
Tech Stack: TypeScript / Javascript / Node.js, ActionHero, React + Next.js, Postgres & Redis, and whole lot of third-party APIs!
-
Launch HN: Hightouch (YC S19) – Sync data from data warehouses to SaaS tools
Congrats on the launch! Hightouch looks great and this need is real. Things seem to be going well, so I don't think I'm taking too much away by mentioning that we have been been working on Grouparoo, an open source alternative that solves similar pain points.
A few differences: git developer workflow focused (branches, CI, PRs, etc), ability to self host, segmentation in destinations (tagging people in mailchimp based on rules, for example)
https://www.grouparoo.com
-
Reverse ETL
We are building Grouparoo. Obviously, it's a biased sample but we are seeing a few trends at play.
-
What software or coding tools are you trying to get your company to invest in?
Has anyone heard or used of reverse etl or Hightouch or open sourced Grouparoo?
What are some alternatives?
webviz - web-based visualization libraries
rotki - A portfolio tracking, analytics, accounting and management application that protects your privacy
haystack - :mag: LLM orchestration framework to build customizable, production-ready LLM applications. Connect components (models, vector DBs, file converters) to pipelines or agents that can interact with your data. With advanced retrieval methods, it's best suited for building RAG, question answering, semantic search or conversational agent chatbots.
airbyte - The leading data integration platform for ETL / ELT data pipelines from APIs, databases & files to data warehouses, data lakes & data lakehouses. Both self-hosted and Cloud-hosted.
metamask-extension - :globe_with_meridians: :electric_plug: The MetaMask browser extension enables browsing Ethereum blockchain enabled websites
TileDB - The Universal Storage Engine
aioli - Framework for building fast genomics web tools with WebAssembly and WebWorkers
meltano
Baserow - Open source no-code database and Airtable alternative. Create your own online database without technical experience. Performant with high volumes of data, can be self hosted and supports plugins
streamlit - Streamlit — A faster way to build and share data apps.
FrameworkBenchmarks - Source for the TechEmpower Framework Benchmarks project
PostHog - 🦔 PostHog provides open-source product analytics, session recording, feature flagging and A/B testing that you can self-host.