I think this is the first time I've seen a model not know which model it was.
217 karma · joined July 13, 2011
I think this is the first time I've seen a model not know which model it was.
What you're seeing is not uncommon in America, especially in the mountains. Roads for traffic in the opposite direction can be separated from your road.
It's also often the case that the other road might be on the other side of a hill/mountain. I've seen that in multiple states across the US.
This specific instance, where two separate roads are going in the same direction can occur where the left road doesn't have exits/entrances, and can move faster than the right road when there's traffic. The left road may also be switched to change traffic directions. I most recently saw this in some mountainous terrain in Arizona.
One thing to think about, which I also struggle with when it comes to large and complicated datasets, is the UI. Even being in the search industry for a long time, it's difficult for me to concretely see how I would use this.
I'd suggest taking a small sample of the dataset that might be reflective of how people would use it, then make that segment public and immediately searchable without registering. eg: One year of articles related to the Olympics.
What I've found is that it's hard for a lot of people to imagine how they would use something without actually using it. So giving people the actual experience of searching the archive and interacting with the results would go a long way.
Again, congrats on the work. This is really impressive work.
Opus 4.7 is available today for 7.5 credits per prompt.
They have also suspended new signups.
After testing all of the major IDEs/tools that integrate with LLMs over the last four weeks, I was happy to settle on Copilot. I, and others, seem to be a lot confident in that decision. Especially since there seems to be no refund path for people who prepaid for a year.
In my 30+ years online, I've never seen an industry change so much in terms of pricing, service levels, etc, as I have the last two months.
I'm really curious where all of this lands, and if AI coding tools will be something that only a small percentage can genuinely afford at a competitive level.
Edit: I just looked around for your YOShInOn RSS reader code and couldn't find it. I did find a number of references it looks like you've made to it on various forums, etc over the years.
That could be something mundane, but I'd like to believe something crazy happens if you yell at it [1]...
[0] https://www.brendangregg.com/blog/images/2025/brendanoffice2...
Later, adding things like analytics and tracking (eg: not just in social media, but also in email campaigns) became another reason to use them, especially for those less tech inclined.
There are four so far. Not sure if there will be more: https://www.pbs.org/show/pompeii-the-new-dig/
This is a minor example, but since you asked...
https://github.com/Cyan4973/xxHash/blob/dev/xxhash.h#L6432
That's an example of a fair number of accumulators that are stored as XXHash goes through its input buffer.
Many modern hash functions store more state/accumulators than they used to. Previous generations of hash functions would often just have one or two accumulators and run through the data. Many modern hash functions might even store multiple wider SIMD variables for better mixing.
And if you're storing enough state that it doesn't fit in your registers, the CPU will put it into the data cache.
It's also worth asking what rule it would need in order to follow the rule. On occasion, a rule I've added isn't quite followed. So I'll respond immediately pointing out what it did, that the rule is in the file, and then will ask it to tell me how I should modify, or add to, the rule in order for it to be easier to follow.
I'd imagine Claude Code has something similar that might be worth looking into.
I can't remember all the little things that happen, which wouldn't happen in a plain text editor, but if you type hyphen-space, then hit enter, the line is deleted and your cursor stays on that line instead of advancing to the next.
It's a trivial example, but things like that happen.
Instead, there is always either markdown or rich text formatting involved. And there's no ability to disable that.
That always seemed odd to me to force that kind of decision on users.
There's also a SQLite DB available to download of the top 1k tag+attr+value combinations. [2]
[1] https://webparsing.io/blog/hidden-in-html-parsing-page-layou... [2] https://webparsing.io/data/commoncrawl-2024-11-html-tags-att...
"I was able to achive a random WARC file compression size of 793,764,785 bytes vs Gzip's compressed size of 959,016,011" [0]
In hindsight, I could have written that up and tested it better, but it's at least something.
[0] https://github.com/benwills/proposal-warc-to-zstandard?tab=r...
The post is interactive, allowing you to search on the 500 most common values per tag+attribute. There is also a free SQLite database available for download of the top 1,000 values per tag+attribute.
This is the first post of an 8-part series that builds toward writing an article parser, the lessons from which can be transferred to writing any other kind of parser you might want.
This is my first time to publish content like this and I'd love any feedback you might have.
https://plugins.svn.wordpress.org/
If so, I'd imagine creating a mirror of the registry would start there.
Edited to add: Sticking with the context of food, "fasting from food" contradicts someone being a "fast eater."
Just make sure to use a dot between my first and last name.
struct SomeData_Gperf_st {
const char* strKey;
SomeData_et enumVal;
};
%compare-lengths
%compare-strncmp
%enum
%ignore-case
%includes
%readonly-tables
%struct-type
%define slot-name strKey
%define lookup-function-name SomeData_Find
%%
val 1, SOME_DATA__VAL_1
val 2, SOME_DATA__VAL_2
The important option here is the %struct-type option. Using that, it will find and parse the struct definition (above in the configuration), and use that to create the objects in the lookup table.The lookup table is defined after the two percent signs; %%. You simply use comma separated values. And then you just have to define the const char* to use for the lookup string, as defined by `%define slot-name`
So, gperf will then generate a function, SomeData_Find(), which returns a pointer to a SomeData_Gperf_st object. Or NULL if it's not found.
There are a number of other options available, but what you're seeing above are the options I use most often.
And, finally, I execute a command like this to run it with that config:
gperf SomeData.gperf.confg --output-file=SomeData_Find.c --multiple-iterations=100"The major drawback of gperf is that it generates function "exists", while we need a "lookup"."
gperf allows for returning actual values. I use it in a whole lot of ways (enough that I've written code that imports data to then generate enums, execute gperf and import generated code, etc), including for generating lookups for HTTP Verbs.
Let me know if you want to know how that's done and I can give more details.
I'll send you an email later this weekend to connect.