HNHacker News
TopNewBestAskShowJobs

ashkankiani

425 karma · joined October 18, 2019

submissionscomments
ashkankiani··on Coding is not solved
Yeah, this is the only argument for why coding isn't solved I need. If it was, then Anthropic, with their access to infinite money and the best unreleased frontier model, should be able to make the absolute best terminal UI imaginable with the most unbelievable performance, 100% edge case testing, and no bugs ever. Terminal UIs as a concept which have existed for about 50 years now. They should be one of the most well understood things in all of programming.

And yet, here we are.

ashkankiani··on America.gov
Maybe I'm shallow, but I'm going to be suspicious of what "epsteingpt" says that I shouldn't be negative about.
ashkankiani··on How to keep enjoying programming in a world of LLMs
Are you only alive to be productive?
ashkankiani··on Early rogue AI agent activity and attempts to hack found on urlquery.net
If Jensen actually cared and believed that their recklessness was a liability to the public and therefore his own fiduciary responsibility to investors then NVIDIA’s dealing with OpenAI would’ve been materially affected. And they weren’t.
ashkankiani··on Stripe's Knowledge AI Platform
I think the specific cases where it failed was when the customer was one of their earlier ones. My guess is that some data schema migration happened, and older data was being served which didn't quite match. So perhaps recency wouldn't potentially even surface those problems either.
ashkankiani··on Stripe's Knowledge AI Platform
That's funny, because I recall in 2015 hitting their endpoint for a large-ish customer, and if you added a boolean to get the total result count, it would 500 every time, presumably because it was doing some kind of "SELECT count(1)" over a postgres table. IIRC, Stripe was ruby internally for a long time, no?

Also their documentation was frequently just straight up incorrect (as in the described json schema for a response was violated. keys missing, different field names, etc.).

But it's been over 10 years, has it improved since then? I'm still in my impression of their stack from back then, although they were decently mature by then as well.

ashkankiani··on Woman Arrested, Dragged Away After Speaking About Flock at City Council Meeting
You might want to look at how a small mob orchaestrated by and comprised of Bush's campaign flooded voting offices to stop a recount, as part of a successful effort to steal the election from Gore https://en.wikipedia.org/wiki/Brooks_Brothers_riot

They didn't even need torches to subvert democracy. I think a majority of Americans are under the misconception that they live in a civilized democracy, and that they have been for a while. But if you look under the covers (or if you're a minority), or just learn history, then you'll find that it's not.

ashkankiani··on Apple has added persistent 'ads' to iOS, and it's driving users crazy
In the US, iPhone 14 Pro has dual eSIM, in Hong Kong, it's 2 physical SIM. I'm not sure why they do this.
ashkankiani··on Apple has added persistent 'ads' to iOS, and it's driving users crazy
I got a third party replacement battery in my iPhone. The battery works just fine and the reason is simple, I no longer live in the US, where the phone was purchased, and if I want to keep my 2 eSIM slots, then I can't get the phone fixed in Hong Kong. Anyone who thinks Apple is very convenient has never lived in multiple countries, imo.

But that's not the worst of it. It took me some time to realize this, but when I wanted to see why my battery was draining so fast one day (I suspect Youtube has a long standing bug that causes it to go turbo drain mode for some reason, the app is incredibly poorly programmed and leaks memory over time, if you didn't know).

But lo and behold, I couldn't see the per-app battery usage breakdown! Why is that? Turns out Apple disables it if it can't detect a third party battery. Why? I'm sure someone will come and try to defend it, but if you do, you better come with an EE degree and math. To me, it looks like a punishment for not being able to go back to USA within my warranty period to replace my phone's battery which was down to 75% capacity.

ashkankiani··on Writing Rust code that's fast by asking agents to make the code faster
I think they can be useful for quickly iterating through benchmarks and trying lots of ideas, but they won't come up with them on their own. Also, I'm not sure why, maybe some mean reversion thing, but they will never, ever suggest writing a tool to make their own life easier, get more accurate information, or anything. Once I point it at a tool, it can be ok at using it (I say ok because they seem to skim the help docs, which is truly ironic, considering I seem to read it more thoroughly even though I'm 100x slower at it. I assume this is some token saving system prompt), but they won't suggest it for you.

This is why I'm not worried about being replaced for now or the forseeable future. For all of the improvements they've made, this part just never seems to change. They could slap another heuristic prompt for the edge case, but eventually it'll revert to the mean again.

I think there is a way to use LLMs to help with programming, but not when I'm not the driver in the seat writing the tests and deciding the architecture. Also I would never ship code written by them as the final product for anything I care about. Since I, like most people, find reading code to be arduous. The more fun thing to do is to force yourself to rewrite it all, treating the LLM's work as a rough draft.

ashkankiani··on Can gzip be a language model?
I'm guessing this is an allusion to reinventing something that already exists, but do you mind explaining what that is to me, since I don't know?
ashkankiani··on I built non-autoregressive decision models with RL a year ago
People have been having this same debate in a very similar way on typed languages vs untyped interpreted languages. I think that, in a similar vein, if you look at the trend over time:

- the addition and standardization (with incomplete coverage) of the solution of adding typing to Python

- how much people are re-discovering the value of performance + typing (e.g. Rust)

then I'm going to take a small leap and extrapolate that the trend will be similar here.

The equivalent of the "one off script in python" will be the LLM, and the long term stable and maintainable solution will be something much more structured and focused like Jev.

ashkankiani··on Android 17 is the first since 3.x to add new APIs without releasing to the AOSP
And if I asked people to name companies which retain a good public perception of being not entirely profit driven, the ones that would most likely come to mind are ones like Costco, where there's someone who is setting the tone and not leaving it to the committee of free market/investors to decide the fate of the company. I wonder if there are parallels here to other similar economic models, hmm...
ashkankiani··on Why building a Rust LSP is hard
This is a good observation, with the observation that there are two different levels of latency requirements. Most of the time, I would be pretty happy with an untyped, unsemantic, mildly smart heuristic based ident completion while the asynchronous semantic one finishes.

The nice thing about this dual setup is that I tend to only want the semantic one if I'm thinking more, so there's naturally a larger time budget for it.

One must always think about the experience they want in UX first, rather than the tools they want to build.

ashkankiani··on Show HN: Scry, programmable internet search w/ congestion pricing
Cool idea. The pricing model reminds me of my time working in algorithmic trading, haha. I'll try this out for some queries I wanted to run.

I suppose the scraping you're doing is a huge part of your value proposition, but I would like to gently nudge you in the direction of making the datasets available via p2p (e.g. a torrent) like how Wikipedia distributes its snapshots in the spirit of democratizing access to data that is becoming increasingly walled off. Also, I think another potential benefit that kind of bulk sharing would have is relieving the congestion from those doing the equivalent operation to extract data via the querying interface.

ashkankiani··on Neovim have a ~$800k Bitcoin donation sitting untouched since 2023
Yes, sorry, I should've said "first version of the built-in lua lsp." I basically laid down the foundations of a house and you guys built the rest of it, haha. No one wants to live in just the foundations except for the brave ones who can imagine the future shape of the house.

I'm slowly starting to come out of my shell again.

ashkankiani··on I Don't Like LLMs
I agree in sentiment, and really a lot of this is Google's doing. It's been around a decade since they switched to approximate results and started encouraging people to search via sentences. I believe this is the one? https://en.wikipedia.org/wiki/Google_Hummingbird

The internet for all of its activity is becoming harder to index because the classic dumb search engines were turned into personalized semantic people pleasers.

The classic model for knowledge discovery still exists, though. You find a blog you like, and use that as a thread of knowledge. You join a community and talk with someone and share resources together. Join a small group chat and post with those people. Your knowledge sharing comes from people you share interests with.

Go to a library, ask your coworkers. If you put in the ground-work, it's possible to do that. For seeds of knowledge outside of your local network, at this point you have to find the indexers that work for you.

Personally, I am in the process of writing the infrastructure for my own search engine, and maybe I will share the tech broadly one day. For now, I'll keep my inventions to my closer personal circles.

In general, I think people should be able to more easily maintain their own offline indices, and share knowledge graphs peer-to-peer rather than relying on a centralized one, in my opinion. This tech has no monetization opportunity, though, so I'm guessing that's why it's relatively under-developed, but luckily for me I have enough money and now enough time to develop and release it (and similar tools).

ashkankiani··on I Don't Like LLMs
He says this in the post almost word-for-word. Are you agreeing with him? The opening "Well" confuses me on what tone you're going with.
ashkankiani··on Neovim have a ~$800k Bitcoin donation sitting untouched since 2023
For a reference point, when I was unemployed and between jobs, I got involved with Neovim for fun, and after some contributions, I eventually tried some full(ish)-time paid work. I wrote the native lua LSP client (:h vim.lsp, and the nvim-lspconfig repo) for Neovim in a few weeks for about $3k USD (circa 2019)? This amount could fund quite a lot of work to be sure.
ashkankiani··on Japan's book scene is moving from bookstores to libraries
If you saw a homeless person in a library, would you be able to tell them they smell to their face? Do you have that kind of "bravery"? Or is it just here, amongst your "peers."
ashkankiani··on AWS says it can't restore some data from mideast facilities struck by Iran
I would love if someone could tell me the name for this phenomena. FWIW, using an LLM incites this same undue assurance as well. No matter how much you know that an LLM might just be hallucinating, it still happens anyway. It's a hard instinct to fight.
ashkankiani··on Navier-Stokes – Tristan Buckmaster [pdf]
The malice would be in not prioritizing the provenance tool at the start as a requirement of the rest of the product. Ethics would tell you that if you can't make the product in an ethical way, then you probably shouldn't make it.
ashkankiani··on Navier-Stokes – Tristan Buckmaster [pdf]
It feels convenient to not spend time on engineering around tooling that could be used to answer a question like “did you violate copyright by training on X?”
ashkankiani··on 216M Spy TVs – The LG Smart TV Problem [video]
In China, CEOs go to prison or are executed. Not all the time, but enough. In the US, they are given a golden parachute and make more money at their next posting. There is mostly only failing upwards.
ashkankiani··on How accurate have Ed Zitron's AI skeptic predictions been?
If you watch his latest interviews you can see him talking about how he finds it useful for simple cases like in the Bloomberg terminal for generating code for queries, but it still has fundamental limitations for doing anything autonomously.

As someone who has used LLMs heavily at work and at home in various serious experiments, I agree. It still requires heavy babysitting and a lot of its limitations wrt context length are fundamental, not something that’s going to be easy to overcome.

ashkankiani··on The creator of Jujutsu has joined ERSC
I pretty much only use jj now, and it was really good even just with the CLI. But adding jjui changed the story entirely. I now rarely use jj without just booting up jjui first.

It will be good to see what jj will look like with more funded dev work, but I'm always a little worried about financial incentives mixing with the tools I use for the long term. I guess the saving grace is that I don't really need more upgrades to jj or jjui as it stands. I'm pretty much content with the features and so I could just save this copy of the repo for future reference.

As far as large assets goes, I have my own VCS-ish system which I just integrate with jj, but it would be nice to see a non-git backend handle large assets better as well.

ashkankiani··on Claude Fable 5.1 and Claude Mythos 5.1
I canceled my Claude subscription, though I did get some utility out of it, because of how much steering was required to use it on complex projects.

A big reason being that anyone who is using Fable seriously will run out of usage limits very quickly, and so will lean on the "Fable for review + design discussion, Opus 5 agents for implementation" paradigm. But an incredibly annoying UX problem is that the resulting report from the agents that Fable reads isn't surfaced to us in the main dialog, it's only summarized back to us (unless you idle in the agent's window to avoid it closing so you can read what it said directly). As a consequence of this game of telephone, the Fable agent will start using some "terms of art" that it and the agents invented, leaving out literally all context that would be useful in helping me understand what converged/diverged from the implementation attempt. It will often try to ask me for input or say that I have to deliberate on something while also referring to things I've never seen (from the agent result) and without providing any context.

I have to repeatedly prompt it to verbosely explain every time (putting it into the system prompt did little to improve this) and remind it that I can't see what the hell it's talking about.

I'm not sure I'll re-subscribe or even really use AI again because it's honestly more frustrating than it's worth, and so the net emotion I'm left with is frustration and without the satisfaction of learning + building something myself. But at the very least, I thought I'd give someone at the company a tip on what seems to me like a common and obvious UX/UI/workflow failing for using Fable, as some last bit of good will.

ashkankiani··on How accurate have Ed Zitron's AI skeptic predictions been?
This isn't a serious analysis of the actual thesis or any of the predictions in any meaningful way. I learned almost nothing from this shallow wall of text. It seems overly focused on "predictions" instead of the ideas behind the predictions, i.e.

- the existence and severity of the financial AI bubble

- the claimed efficacy of AI in terms of its utility vs the actual observed utility

- whether the net good provided by AI outweighs its very heavy costs

I find the section of listing a bunch of selected "predictions" and just saying "Wrong" to elucidate very little. Not that a sentence is sufficient to provide explanation, but Dan stops even doing that bare minimum partway through and just saying "Wrong" full-stop. The reasoning is left up to the reader I guess?

How is it wrong? What was the actual thesis behind it? Is the underlying idea wrong or just the specifics on execution? Was there undetermined factors that mled to the wrong prediction? What can we learn from those factors in order to update our model?

We saw that even though the underlying financials in 2008 were trash and lots of people knew they were trash, things didn't quite collapse in the time frame or way we expected, because an unknown part is how much shenanigans companies can do to extend the runway.

As an example, credit ratings agencies didn't drop ratings to match reality because of customer relation incentives, which is a factor that is not easy to account for and strongly affects the timing of the collapse.

I find the positive reactions to this blog to be confusing. I feel like I learned nothing at all, which makes sense considering under "why write this?" he says "I got four hours of sleep and my brain wasn't good for much of anything and I saw someone posted a screenshot of a reddit post dunking on Ed Zitron's prediction record."

ashkankiani··on Quack: The DuckDB Client-Server Protocol
My first thought: setting up a self replicating duckdb wrapper over ssh so that I can execute queries on any computer. Can’t wait to play with this!
ashkankiani··on setBigTimeout
You are a bad programmer if you think silently doing the wrong thing is not a bug. The right thing to do with unexpected input as the setTimeout library author is to raise an exception.
Page 1 of 5Next →