990 karma · joined May 26, 2015
Most recently: Salesforce via Twin Prime acq. Before that Xilinx. Interned at CCT@LSU, Computer Lab in Cambridge UK, Ericsson, HIIT.
WWW: http://www.koszek.com Twitter: http://www.twitter.com/wkoszek
[ my public key: https://keybase.io/wkoszek; my proof: https://keybase.io/wkoszek/sigs/wFGIqesxDG511PD1Jve8FJ7qQ0h9Sc6xvGw-NTCddek ]
SELECT COUNT(*) FROM large_text_db WHERE X
Where X is something that must be matched exactly. X can be FTS query on FTS-indexed table, but the way COUNT() works in PG is that it's impossible to make it fast. Over large tables, lets say 1B+ rows, it can be very very slow.
Example use case is: searching through a hospital DB of reports that have "pancreatic cancer" in them. This is trivial in SQLite, but in PG it's hard.
And your sales folks would call and say: "No need to change anything, we still run PostgreSQL, and ours is just called pgrust, but it's N times as fast".
I published similar project here: www.emuko.dev - emulator for RISC-V. This one turned out to be 3x as slow as QEMU for example.
re: CTOs - if the improvement is 3% nobody of course will look at it. If the improvement is 30%, it'll be too big for big players to ignore, so as a CTO you'll be tasked with trying it out. It's really a matter of whether this thing is safe and secure, losing data or has a trojan. If the authors can prove it's all valid working code etc., it'll be a viable project.
I'm building data + AI platform. It got complex, and I'm using AI-assisted coding to move fast. One thing that helps me was property based testing. I have a traffic generator that simulates 10 users working on my platform. I run it 24/7 and if it shows that the software survives the test (I called it "fate" after ffmpeg's CI), it's good enough to roll out. If you wrote something like that, core PostgreSQL folks would like it too, unless there's something equivalent like this. It'd be: create random tables, fill with random data, then issue randomly constructed query.
Another hope I have that we could learn math/physics that way. Having a better more visual and intuitive understanding of some math concepts would be great. Perhaps having a truck go through the plot of a function and learning about limits and maxima and minima would be great etc. Same story: playing something like this for several hours could burn in kid's/adult's mind a pattern of undersatnding calculus that would last a lifetime.
We already have patients trying to track their own health over longer time which is great. We then just have to make AI good enough to spot warning signs (without patients asking). Or parhaps we need to make those tests easy and cheap and regular.
The faster and earlier we start to scan everyone regularly, as long as scanning methods aren't invasive, the more certainty we'll have what to warn people about and what not to tell them. Perhaps with the regular screening (imaging quarterly, if the scan is fast) you could see what is growing and what isn't.
https://www.nextbigfuture.com/2022/02/spacex-reusable-rocket... -- looks like target price for Starship launch would be $3--$5m according to the author.
Wouldn't the /kg price to SpaceX be:
3000000/100000 = $30/kg -- 5000000/100000 = $50/kg?
If they recover everything and produce fuel at scale, wouldn't it drop the cost even more.
What many people quote here are commercial rates, I think. SpaceX won't pay those prices.
Can someone check my math
You can give it a shot in VM - I've done testing on macOS, and then used Ubuntu 24.04 home box to install it, also in multipass. In GitHub I put howto.