Indeed, the source link is about exactly this: a crappy scan appears to have 彁 but a better copy reveals it was 彊.
272 karma · joined November 30, 2020
otherdavidli@lidavidm.me
Indeed, the source link is about exactly this: a crappy scan appears to have 彁 but a better copy reveals it was 彊.
Or the author's newsletter: https://www.construction-physics.com/
https://en.wikipedia.org/wiki/Pixel_shift
EDIT: Sigma also has "Foveon" sensors that do not have the filter and instead stacks multiple sensors (for different wavelengths) at each pixel.
Could you also get d2 into GitHub and Notion while you have it here :)
Having to figure out the exact pixel widths defeats the point of these tools, at least for me.
Which links to this blog post explaining the choice in more detail: https://smallcultfollowing.com/babysteps/blog/2022/04/12/imp...
[1]: https://bookshop.org/info/ebooks ("Can I read my ebooks on my Kindle, Kobo, Nook, etc.?")
You can sort of see what benefits you might get from a post like this, though: https://willayd.com/leveraging-the-adbc-driver-in-analytics-...
While we're not using Arrow on the wire here, the ADBC driver uses Postgres's binary format (which is still row oriented) + COPY and can get significant speedups compared to other Postgres drivers.
The other thing might be to consider whether you can just dump to Parquet files or something like that and bypass the database entirely (maybe using Iceberg as well).
Arrow Flight SQL defines defines a full protocol designed to support JDBC/ODBC-like APIs but using columnar, Arrow data transfer for performance (why take your data and transpose it twice?)
https://arrow.apache.org/blog/2022/02/16/introducing-arrow-f...
There's an Apache-licensed JDBC driver that talks the Flight SQL protocol (i.e. it's a driver for _any_ server that implements the protocol): https://arrow.apache.org/blog/2022/11/01/arrow-flight-sql-jd...
(There's also an ODBC driver, but at the moment it's GPL - the developers are working on upstreaming it and rewriting the GPL bits. And yes, this means that you're still transposing your data, but it turns out that transferring your data in columnar format can still be faster - see https://www.vldb.org/pvldb/vol10/p1022-muehleisen.pdf)
There's an experiment to put Flight SQL in front of PostgreSQL: https://arrow.apache.org/blog/2023/09/13/flight-sql-postgres...
There's also ADBC; where Flight SQL is a generic protocol (akin to TDS or how many projects implement the PostgreSQL wire protocol), ADBC is a generic API (akin to JDBC/ODBC in that it abstracts the protocol layer/database, but it again uses Arrow data): https://arrow.apache.org/blog/2023/01/05/introducing-arrow-a...
I think the difference is more that DataFusion is built as a library so you can plug it into the product you're building (e.g. Comet, which plugs it into Spark, or pg_lakehouse, which plugs it into Postgres). Polars could be used that way, but it's also a functional package you can pip install and use as a Pandas alternative right now.
Coincidentally I was looking into C++ documentation generators again.
In terms of integration, what I've settled on for apache/arrow-adbc is using Sphinx as the toplevel site generator, then writing a script that generates fake Intersphinx indices for a Doxygen site. That way you can link to Doxygen items from within Sphinx without having to hardcode URLs, instead by referencing a class name or similar, and Sphinx will warn if you reference something nonexistent, without having to use something like breathe that tries to render the Doxygen output within Sphinx. (Same approach with Javadoc -> Sphinx, too.)
https://invent.kde.org/graphics/krita/-/merge_requests/1783 https://krita-artists.org/t/can-we-get-mixbox-on-krita/64201...
You can also buy from third party stores, e.g. Weightless Books for sci-fi & fantasy, and just drag-and-drop the epub onto the Kobo.
https://atadistance.net/2019/10/20/japanese-text-layout-for-...
> Baseline font metrics will never deliver great CJK typography because there are too many limitations. > > This is why InDesign J implements virtual body metrics based on Adobe proprietary table information for true high-end Japanese layout. There is no virtual body standard digital font metric standard so everybody implements the missing stuff on the fly and everybody does it different. Unfortunately the irony of it all is that Adobe played a huge role in how these limitations played out in the evolution of digital fonts, desktop publishing (DTP) and the situation we have today.
I have a Kobo reader which supports both ePub 2 and ePub 3, and IIRC you need ePub 3 in order to get proper RTL/top-to-bottom text and Japanese typesetting, as well as proper comics support (if you buy an ePub 3 manga, it'll properly flip the page turn direction and the progress bar; a CBZ or other format won't). But most other readers I run into don't understand ePub 3 properly.
Anyways, I think I could have dealt with it if it handled large books fine.
(Also, a lot of manga gets distributed through proprietary apps now so an iPad is probably your best bet anyways, at least if you read the serialized version and not the tankoubon releases...)