usv
goawk
Stars - the number of stars that a project has on GitHub. Growth - month over month growth in stars.
Activity is a relative number indicating how actively a project is being developed. Recent commits have higher weight than older ones.
For example, an activity of 9.0 indicates that a project is amongst the top 10% of the most actively developed projects that we are tracking.
usv
- How to fix CSV: make it even more U+1F4A9 PILE OF POO
-
Friends don't let friends export to CSV
The reason why USV did not use the proper ASCII codes for field separator and record separator is a bit too pragmatic for me…
https://github.com/SixArm/usv/tree/main/doc/faq#why-use-cont...
-
Show HN: Comma Separated Values (CSV) to Unicode Separated Values (USV)
https://github.com/SixArm/usv/blob/main/doc/criticisms/index...
-
Ask HN: Can you help with IANA, + RFC and ABNF?
I'm working on standardizing a data exchange format and media time, and seeking advice here from anyone experienced with how to do this.
For example, how to properly submit a request to Internet Assigned Numbers Authority (IANA.org), or Internet Engineering Task Force (IETF.org), and proofread the Augmented Backus-Naur Form (ABNF).
The project is Unicode Separated Values (USV) and media type is "text/usv".
Work in progress here: https://github.com/sixarm/usv
My email is [email protected]
Thanks!
- Show HN: Unicode Separated Values (USV) for data formatting
-
Modernizing AWK, a 45-year old language, by adding CSV support
Ben this is great, thank you. Would you be open to adding Unicode Separated Values (USV) as well? It's much like CSV and also simpler because of no escaping and no quoting. I can donate $50 to you or your charity of choice as a token of thanks and encouragement.
https://github.com/sixarm/usv
- GitHub - SixArm/usv: USV: Unicode Separated Values
- Show HN: USV = Unicode Separated Values
goawk
- GoAWK, an Awk interpreter written in Go (2018)
-
The Awk Programming Language, Second Edition
TIL: GoAWK [1] - A POSIX-compliant AWK interpreter written in Go, with CSV support.
[1]: https://github.com/benhoyt/goawk
- Looking for a script for csv file
-
Anyone else doing compiler work in Golang?
Another nice project that I have used from time to time (and a very good source for insight) is the awk interpreter written in go https://github.com/benhoyt/goawk
-
Tool to interact with CSV
No, I want exactly the opposite - it should be a , b,c as a single string field containing a literal comma, and c. For example, https://github.com/benhoyt/goawk has csv support. https://github.com/benhoyt/goawk/blob/master/docs/csv.md - more info.
-
Why does awk parse '1&&x=1' as '1&&(x=1)' not '(1&&x)=1' when '&&' is high precedence than '='?
I've had a go at solving this in this PR -- feedback welcome. I don't love it, but oh well, it solves the problem at hand. Your comment pointed me in the right direction, thanks again.
-
Looking for programming languages created with Go
There are quite a few re-implementations of scripting languages like Lua in Go. I've written an AWK interpreter in Go.
-
Oracle DB support in Benthos
github.com/benhoyt/goawk -> this library lets you embed an AWK runtime in your applications, very easy to use and useful for enabling some powerful scripting in things you build
-
Brian Kernighan adds Unicode support to Awk (May, 2022)
Yes, that's right. With my simplistic UTF-8-based implementation it turned length() -- for example -- from O(1) to O(N), turning O(N) algorithms which use length() into O(N^2). See this issue: https://github.com/benhoyt/goawk/issues/93
Similar with substr() and other string functions, which when operating as bytes are O(1), but become O(N) when trying to count the number of codepoints as UTF-8.
GNU Gawk has a fancier approach, which stores strings as UTF-8 as long as it can, but converts to UTF-32 if it needs to (eg: the string is non-ASCII and you call substr).
It looks like Brian Kernighan's code has the same issue with length() and substr(). I'm going to try to email him about this, as I think it's kind of a performance blocker.
-
Ask HN: Is having a Personal blog/brand worth it for you?
I'm not sure if it was via my personal website or just my GitHub profile, but I got my current job at Canonical due to the CTO there reaching out about my GoAWK project (https://github.com/benhoyt/goawk). I get regular recruitment emails because I have my CV/resume online: most of them are very low-effort, but 1 in 20 or something are interesting emails where the recruiter has actually looked at my website and will tailor it personally. I also just enjoy technical writing, and get joy out of sharing it on HN. So it's "worth it" for me.
What are some alternatives?
tsv-utils - eBay's TSV Utilities: Command line tools for large, tabular data files. Filtering, statistics, sampling, joins and more.
bytehound - A memory profiler for Linux.
qsv - CSVs sliced, diced & analyzed.
nio - Low Overhead Numerical/Native IO library & tools
awka - Revive awka - Awk to C Compiler
csvquote
intellij-awk - The missing IntelliJ IDEA language support plugin for AWK
tumblelog - A static tumblelog generator available as both a Perl and Python version
awk - One true awk
bashcc - C compiler written in Bash script
tectonic - A modernized, complete, self-contained TeX/LaTeX engine, powered by XeTeX and TeXLive.