Sep
geoparquet
Sep | geoparquet | |
---|---|---|
2 | 3 | |
639 | 725 | |
- | 4.4% | |
8.8 | 5.5 | |
12 days ago | 4 days ago | |
C# | Python | |
MIT License | Apache License 2.0 |
Stars - the number of stars that a project has on GitHub. Growth - month over month growth in stars.
Activity is a relative number indicating how actively a project is being developed. Recent commits have higher weight than older ones.
For example, an activity of 9.0 indicates that a project is amongst the top 10% of the most actively developed projects that we are tracking.
Sep
-
Eli Bendersky: Faster XML Stream Processing in Go
Pretty sure the C performance would have been trivially matched by C# were it to be ported there.
While not XML (which does not see much love from the community to surprise of HN), gigabytes per second of parsing throughput are reached for CSV without sacrificing UX: https://github.com/nietras/Sep?tab=readme-ov-file#packageass...
-
Friends don't let friends export to CSV
If you ever need to parse CSV really fast and happen to know C#, there is an incredible vectorized parser for that: https://github.com/nietras/Sep/
geoparquet
-
Friends don't let friends export to CSV
That's why I'm working on the GeoParquet spec [0]! It gives you both compression-by-default and super fast reads and writes! So it's usually as small as gzipped CSV, if not smaller, while being faster to read and write than GeoPackage.
Try using `GeoDataFrame.to_parquet` and `GeoPandas.read_parquet`
[0]: https://github.com/opengeospatial/geoparquet
-
COMTiles (Cloud Optimized Map Tiles) hosted on Amazon S3 and Visualized with MapLibre GL JS
GeoParquet
-
Postgres and Parquet in the Data Lke
> "Generating Parquet"
It is also useful for moving data from Postgres to BigQuery! ( batch load )
https://cloud.google.com/bigquery/docs/loading-data-cloud-st...
Thanks for the "ogr2ogr" trick! :-)
I hope the next blog post will be about GeoParquet and storing complex geometries in parquet format :-)
https://github.com/opengeospatial/geoparquet
What are some alternatives?
mbtiles-spec - specification documents for the MBTiles tileset format
odbc2parquet - A command line tool to query an ODBC data source and write the result into a parquet file.
geemap - A Python package for interactive geospatial analysis and visualization with Google Earth Engine.
flatgeobuf - A performant binary encoding for geographic data based on flatbuffers
postgres_vectorization_test - Vectorized executor to speed up PostgreSQL
BlenderGIS - Blender addons to make the bridge between Blender and geographic data
parquet_fdw - Parquet foreign data wrapper for PostgreSQL
com-tiles - Streamable and read optimized file archive for hosting map tiles at global scale on a cloud object storage
zarr-python - An implementation of chunked, compressed, N-dimensional arrays for Python.
tippecanoe - Build vector tilesets from large collections of GeoJSON features.
tilemaker - Make OpenStreetMap vector tiles without the stack
duckdb_fdw - DuckDB Foreign Data Wrapper for PostgreSQL