HNHacker News
TopNewBestAskShowJobs

oddthink

481 karma · joined June 22, 2010

submissionscomments
oddthink··on What is Glamorous Toolkit v1.0?
On Mac at least, you just have to download the zipped disk image and copy over the contents. I just put it in ~/Applications, and it worked.
oddthink··on Mojo is now available on Mac
Naive question, but how do they distinguish themselves from Julia, which is also in that space?
oddthink··on Every app that adds AI looks like this
I've found LLMs to be very useful for, well, text-based things. I've not found the "bot" implementations useful, but they're better tech for summarization, highlighting important sections (i.e. what sentence of this product review should I show in bold, given the search term), and entity recognition (what are all the products mentioned here).

They are expensive to run, in terms of GPU cycles, but they are noticeably better than the previous models.

It's also hard to constrain them well. If you want 95% accuracy, it takes some tuning work. If you also want to avoid 1% total batshit nonsense (repeat "chicken" 50 times), then you have to check for that. Earlier models were sometimes wrong, but they were not quite so aggressively wrong as the 1% case of LLMs.

That's just my anecdotal experience, but it leaves me both optimistic about applications in the right spaces and worried that people are just shipping something that's OK 75% of the time and calling it a product.

oddthink··on Graph Mining Library
Clustering. I used the correlation clusterer from here for a problem that I could represent as a graph of nodes with similarity measures (this data looks like this other data) and strong repelling features (this data is known to be different from this other, so never merge them).
oddthink··on Llama.cpp: Full CUDA GPU Acceleration
It was fine, back in the early days (I started with 1.4-ish). I just downloaded the tarball, unpacked, configured, make, installed into /usr/local on my workstation, then downloaded and stuck any packages into site-packages. Numeric was sometime tricky to compile right, but ye olde "configure && make && make install" worked fine.

Of course, that worked because 1) I was really only doing one project, not juggling multiple ones, 2) there weren't all that many dependencies (Numeric, plotting, etc.), and 3) I was already up to my eyeballs in the build system with SWIG and linking to the actual compute code, so I knew my way around the system.

But every now and then I just shake my fist at the clouds then mutter darkly about just installing the dang thing and maybe not taking on so many dependencies. :-)

oddthink··on How to Finetune GPT-Like Large Language Models on a Custom Dataset
It's worth it whenever you have a reasonable amount of training data. You can get substantial quality improvements automatically. Unless you're doing some kind of prompt-optimization, prompt-tuning is a lot of random guessing and trial-and-error. It's also most necessary when you have a smaller base model, as opposed to one of the big ones.
oddthink··on How to Finetune GPT-Like Large Language Models on a Custom Dataset
Wouldn't a vector database just get you nearest-neighbors on the embeddings? How would that answer a generative or extractive question? I can see it might get you sentiment, but would it help with "tell me all the places that are mentioned in this review"?
oddthink··on A brief interview with Tcl creator John Ousterhout
I do prefer tcl quoting to bash quoting, but you do have some mental overhead when writing procedures, since unpaired braces inside quotes may do surprising things, if you're thinking "writing a command" rather than "raw strings". Comments are similar.

That being said, those are very much edge cases.

More damning, from my POV, is that you can't get ref counting of things like C objects or file handles, since they're just string handles. But there are a lot of uses that don't need that

oddthink··on Turns are better than radians
Oh yes. This is definitely a pet peeve of mine. CGS is so much nicer. I did E&M from Jackson before he converted it to MKS, and I still can't keep all those epsilon_0 and mu_0's straight. (Not that it comes up all that much.)
oddthink··on Developers spend most of their time figuring the system out
More on topic, I've experimented some with Glamorous Toolkit, and I really liked what I saw of it, but (somewhat ironically), I need to dive more into its own model and understand it before I can use it for anything real.

For example, I couldn't find docs on its display stack, so when I wanted to display a graph but with images instead of text for the nodes, I got stuck. Sure, I could dive into the code and eventually get something that worked, but I wouldn't be sure I was coding to the API or to the implementation. It crossed over to "playing around with this tool" to "have to do some real work to understand", so I ended up dropping it.

oddthink··on Developers spend most of their time figuring the system out
I still don't really get pair programming. Maybe it's my inner introvert, but it sounds really draining. And what about all of the non-programming time during the day?

I mean, today I spent time:

- Digging into the analysts dashboards for some time series that seemed off. Generated a plot from my own metrics stash, filed a bug w/ the analysts about the differences. Replied to some doc comments saying that I'd done so.

- Code review of a few changes.

- Write a quick jupyter notebook (well, colab) analyzing a different problem. Basically reading the data for an example into the notebook, writing a bit of code to visualize it, fiddle until I had some sense of what the problem was.

- Decide I should regenerate a dataset with some different processing. Spent maybe an hour doing the code changes and getting them reviewed. Then build the binary, run it to generate a dataset, then launch a few-hour processing job to see if it works better.

- Write some notes on how I should decide if the new version is better than the old version.

- Fire off email to a few people. Asked one person about existing viz tools, updated another about my earlier colab experiments and asking if they know of anyone who's done a similar analysis.

Out of all this, I think pair programming would have maybe been tolerable for a bit in the middle when I was actually making code changes, but everything else was either based on huge amounts of internal context or was completely ephemeral analysis.

I think I must be doing a different kind of thing. :-)

oddthink··on US Senate votes unanimously to make daylight savings time permanent
I'd definitely choose #3. I don't mind switching, but it's what I'm used to. I'd be OK with dropping it, but if we did, I'd want standard time.

Permanent DST makes no sense to me. Maybe it's my astronomy background, but "noon" means something, something that involves the position of the sun and the earth. We quantize that to timezones for coordination, but it doesn't mean it's meaningless.

If we stop switching, fine, but don't mess with noon. Just change your schedule to 8-4 or whatever. Permanent DST seems like wanting everyone to be above average. Or deciding that everyone would be happier if they're taller, so we're shrinking the foot by 10%.

oddthink··on Updating the most influential book of the BASIC era
Wow. These books were, well, foundational to me. I had a blast typing those into my C64. Then at a summer camp (CTY at F&M) I thought I was so very clever for writing a version of Nim where the computer would silently pick to go first or second so that it always won. I don't know why I still remember that.

And then some summer jobs writing dBase III and FoxPro, some actual education, and here I am.

Those books absolutely set the tone and sparked that first interest.

oddthink··on Floating point numbers, and why they suck
I agree. Enough already. Isn't this covered in something like week 2 of CS 101? If you've missed that, maybe the literally decades of "ermagerd, floats!" articles would give you a hint.

As someone who's worked in science and finance (modeling, not accounting), floats work just fine, thank you very much. The modeling/accounting split in finance is a legit point of confusion, though.

oddthink··on The Economics of Broadway Shows
It's definitely true. My kids go to school very close to Broadway, and it seems like something like half the parents are in some Broadway-related jobs, from performing to backstage to office, a lot of them living in Manhattan Plaza.

It's a very hit-or-miss industry, so there will be periods when you're not doing anything, and then you're left with this pool of extremely high-talent people doing small productions, waiting tables, and the like.

oddthink··on TaleSpire is a beautiful way to play pen and paper RPGs online
I guess I'm not the target audience, but I don't see the appeal of skeuomorphism here over something like on-the-fly zone creation tools. Can someone help me understand the market here?

I've know that the first thing dice rollers seem to grow is 3D simulated dice, so clearly the market is there.

But this would be (for me) replacing an assortment of wood blocks, lego minifigs, scribbled-on index cards, and assorted tokens with a whole lot of additional prep time.

oddthink··on Heuristics for Effective Software Development: A continuously evolving list
I can see this as being a problem, like you said, if there isn't communication, but I don't think that demanding a "PBFD" expert is realistic or scalable.

But in my experience the PM is technical enough to understand what's going on. (Can write some SQL to answer questions, possibly ex-engineer themselves, etc.) They're in the same meetings, same email threads, looking at the same set of OKRs, etc. It's part of the engineers' jobs (B/F/D whatever) to communicate their constraints and their ideas (both product and pure-tech) to the PM, and it's part of the PM's job to take those into account when advocating for what should be done.

Similarly, the more the engineers know each others' specialties, the better they can coordinate. It's probably more important for everyone to have "a little bit of product" in them, but that doesn't mean we don't need a product-specialist.

When it's time for quarterly planning, the PM's voice is definitely loud, but they're still just one voice in the room. They're the one accountable for the product, which gives them some leverage, but the other voices are there (TLs, managers, etc.)

Now I can see this going terribly wrong if the scale is off (only one PM for too many engineers), or if communication breaks down (PM scribbles a "design" on a napkin and faxes it over), or if only the PM is consulted for planning. But the problem there is that communication broke down, not that it's bad to have a PM.

oddthink··on Heuristics for Effective Software Development: A continuously evolving list
Why do agile-folks seem to hate product managers so much?

This comes out in #13, #19, and #20 (well depending on exactly what they mean by "managing".) I find it really helps to have someone dedicated to the product side, so they can help with prioritization, generating ideas for what to try next, keeping track of the research, and tracking the launch process and metrics.

We could, I suppose, distribute all of that to the engineers, but there's value in that specialization.

(I generalize to "agile-folks" here because I feel like I've seen this before, but I can't think of exactly where.)

oddthink··on PathQuery, Google's Graph Query Language
I've been using this fairly heavily recently (internal to Google), to the point where I'm thinking of investing in writing an org-babel mode for it. It's a really nice way to structure queries!
oddthink··on It’s hard work to make ordering groceries online so easy
+1, I've been very impressed by FreshDirect. I have a Morton Williams across the street (crowded, bad produce, iffy selection) and a Whole Foods around the block (pretty nice, but long lines and I didn't really want to go in before vaccination), but FreshDirect has just worked.

Their site and app are glacially slow, though. Not sure what's going on there. And for a while they've had some text-entry glitch that reverses text every now and then. But that's just griping.

Also, with kids it's nice having the weekly 3-4 gallons of milk just show up rather than schlepping them home from the bowels of Whole Foods.

oddthink··on Google made it nearly impossible for users to keep their location private
As a Googler, this sounds a little over-sensationalized to me (mostly because of the choice of quotes), but I bet there was some unpleasant sausage-making going on as well.

First of all, the quote from Jack Menzel about figuring out home and work locations sounds 1) about typical for him and 2) exactly correct. If Google is giving you directions and popping up traffic along your route between home and work every day, it's pretty clear what's going on. There's nothing weird or unexpected going on there.

Second, the settings thing. I see this as two issues: one, multiple settings, and two, making the settings hard to find. For the first, both search and maps had their own settings (web whatever vs location/location history), and they didn't talk to each other. I'm sure the relevant VPs talked about relevant VP-things, which probably did not include the config page. Both were sure they had an option to turn location off (for their project), so that box was checked, and done. Yes, someone should have made sure there was a single button, not three, but the org chart was shipped instead.

The hard-to-find thing is harder. I, as a regular schlub engineer, think that sounds pretty sleazy, but I have no idea how true it is. If someone's A/B test said, oh user engagement was down on this arm, I can see that happening. It'd be a failure, but I can see that happening.

I guess at my level, it seems like all the people I'm working with take user privacy really seriously. If someone wants their data deleted, we go through a lot of hoops to make sure it's really gone. Any feature using user-data gets a privacy review and usually ends up requiring pretty strict differential privacy bounds.

This is a little unfair and possibly just ignorant, but my impression is that Google is far better at protecting user location info than the telecoms, who have more complete data from cell-tower triangulation and who are generally willing to sell that data to whoever, and yet they get a lot less attention for it.

oddthink··on iPad Pro M1
What do you find lacking in the iPad Pro? I'm curious, because I have a 12.3" pixel slate that I'm not a fan of, because it's too big to work well as a tablet (cumbersome to hold, etc.), and inconvenient to attach the somewhat-wobbly keyboard to it, so it feels like it works well for neither use-case. I'm sure the iPad software is better, but the size seems like a contradiction.

My daily driver is my 16" MBP, and while I'm thinking of getting a plain-vanilla 10.2" iPad, I can't think of any cases where I'd want a bigger one.

oddthink··on Summary of C/C++ integer rules
+1 for this. I was just bitten by this last week, when I switched from using a custom container where size() was an int to a std::vector where size() is size_t.

The code was check-all-pairs, e.g.

  for (int i = 0; i < container.size() - 1; ++i) {
    for (int j = i + 1; j < container.size(); ++j) {
      stuff(container[i], container[j]);
    }
  }
Which worked just fine for int size, but failed spectacularly for size_t size when size==0.

I totally should have caught that one, but I just couldn't see it until someone else pointed it out. And then it was obvious, like many bugs.

oddthink··on Geometric Algebra (2012)
This gets directly to my main question each time I see geometric algebra show up: how does it fit in with the "normal" notation of differential geometry? What assumptions does it make for a metric, what is the equivalent of parallel transport, Lie brackets, how does it represent gradients (and other things that are naturally 1-forms), etc., etc.?

All the treatments I've seen jump in to manipulation without really going in to the axioms used. That paper seems to do a lot better at fitting it together, so I'll certainly read it, thanks.

oddthink··on Will R Work on Apple Silicon?
Wow, a little hostile here?

My assertion, that R's NaN is not "non-standard", seems upheld by the article. It's a quiet NaN with a payload, which is well-defined by the IEEE 754 standard.

As other posters pointed out, it's relying on a "should" behavior from the spec, which is risky but common. It sounds like disabling the "RunFast" mode cleared up their issues, which seems quite far from it being an "obscenely bad" design decision.

It's not terribly unusual to require IEEE 754 compliance in numerical code, like the usual options for avoiding --ffast-math -style stuff.

oddthink··on Will R Work on Apple Silicon?
It's not a "non-standard NaN". It's just a particular one, out of many possible quiet NaN values. If the Apple silicon isn't propagating the payload of the input NaN value to output, that's a violation of IEEE 754.

(IIUC, that is. It may be something like a "should" not a "must".)

oddthink··on Cyclone Scheme
Was R7RS-large ever ratified? It doesn't look so to me, but it's been so long that I'm not sure I'm looking in the right places.
oddthink··on What's so hard about PDF text extraction?
Interesting! We're working on OCR on menu photos, which has some parallels in structure, but has a much smaller common vocabulary than a dictionary, almost by necessity. :-)

Many menus are also available in PDF form, so we're trying to figure out if it's worth bothering with the PDF itself, or if we should just render to image and thus reduce the problem to the menu-photo one.

oddthink··on What's so hard about PDF text extraction?
It seems like it should be doable to train a two-tower model, or something similar, that simultaneously runs OCR on the image and tries to read through the raw PDF, that should be able to use the PDF to improve the results of the OCR.

Does anyone know of any attempt at this?

Blah blah blah transformer something BERT handwave handwave. I should ask the research folks. :-)

oddthink··on The Road Less Traveled to Fusion Energy?
No, it's used for control. But you should be thinking more like "process control" than "intelligent" control. Or to be more specific, deriving an optimal process control model via pretty hard-core plasma density reconstructions; my understanding is that they track plasma density fluctuations and tweak parameters in real-time, but I would be surprised if the TF model is in the control path.

It seems a bit overly-aggressive to accuse them of lying because some pop-sci journalist heard tensorflow and went straight to "artificial intelligence." And even there, it's an understandable and common hype.

It doesn't take away from the point that they're doing some really interesting and novel computational physics to make this thing work.

← PreviousPage 2 of 7Next →