1,644 karma · joined June 6, 2009
Contact: sdrinf at google's email service
Web: https://sdrinf.com/
What's also new here, is VRAM-context size trade-off: for 25% of it's attention network, they use the regular KV cache for global coherency, but for 75% they use a new KV cache with linear(!!!!) memory-token-context size expansion! which means, eg ~100K token -> 1.5gb VRAM use -meaning for the first time you can do extremely long conversations / document processing with eg a 3060.
Strong, strong recommend.
* This approach is the _most consistent_ with retaining anonymity on the internet, while actually helping parents with their issues. If any age-relevant gatekeeping needs to be made on the internet at all, this is the one I find acceptable.
* this is because the act very specifically does NOT require age _verification_ ie using third-parties to verify whether the claimed age is correct. Rather, it is piggybacking on the baked-in assumption, that parents will set up the device for their kids, indicating on first install what the age/DoB is, then handing over the device -a setting which can, presumably, only be modified with parental consent
* yes, there are edge cases, esp in OSS, and yes, it would be nice to iron those out -but the risk = probability x impact calculus on this is very very low.
* If retaining anonymity on the internet is of value to you, don't let the perfect be the enemy of good enough.
* even if an openweight model appears on huggingface today, exceeding SOTA, given my extensive experience with a wide variety of model sizes, I would find it highly surprising the "99% of use cases" could be expressed in <100B model.
* Meanwhile: I pulled claude to look into consumer GPU VRAM growth rates, median consumer VRAM went 1-2GB @ 2015 to ~8GB @ 2026, rougly doubles every 5 years; top-end isn't much better, just ahead 2 cycles.
* Putting aside current ram sourcing issues, it seems very unlikely even high-end prosumers will routinely have >100GB VRAM (=ability to run quantized SOTA 100b model) before ~2035-2040.
The purely code part you described is a bit of an "extra steps" -you can just... vscode open target repo, "claude what does this do, how does it do it, spec it out for me" then paste into claude code for your repo "okay claude implement this". This sidesteps the security issue, the deadly trifecta, and the accumulation of unused cruft.
Discord has a financially and politically vulnerable posture that is downstream of having to operate a very large team, raise funding, be exposed to investor market pressure. However, it is also one of the rare instances of successful consumer freemium subscription monetization. A clone does not have to pay the tuition of "what makes this specific space compelling, and want-to-pay-for"; it just have to _exists_, passively soaking up migrants from each platform shift.
ITT WTB 3rd place for my frens.
This is exactly what chatgpt 5 was about. By tweaking both the model selector (thinking/non-thinking), and using a significantly sparser thinking model (capping max spend per conversation turn), they massively controlled costs, but did so at the expense of intelligence, responsiveness, curiosity, skills, and all the things I've valued in O3. This was the point I dumped openai, and went with claude.
This business model issue is a subtle one, but a key reason why advertisement revenue model is not compatible (or competitive!) with "getting the best mental tools" -margin-maximization selects against businesses optimizing for intelligence.
Tolerating this is very bad form from openrouter, as they default-select lowest price -meaning people who just jump into using openrouter and do not know about this fuckery get facepalm'd by perceived model quality.
* Use it as a "source": chatgpt -> settings -> apps & connectors -> add it as your connector. This supports only 2 functions: search, and fetch; details: https://help.openai.com/en/articles/11487775-connectors-in-c... ; in business / edu version there is support for "full MCP mode": https://help.openai.com/en/articles/12584461-developer-mode-...
* Enable "developer mode" chatgpt -> settings -> apps & connectors -> advanced settings -> developer mode. Available on paid&pro levels only. This can do full MCP access, but can't (currently) use your memory settings.
The option that works under all conditions is to use the API, and add it as a function directly (no MCP) -this works regardless what plan you have on openai.
Specific repro steps: set system prompt to: "Current date: 2025-09-28 Knowledge cut-off date: end of January 2025"
Then re-run all your tests through the API, eg "What happened at the 2024 Paris Olympics opening ceremony that caused controversy? Also, who won the 2024 US presidential election?" -> correct answers on opus / 4.0, incorrect answers on 3.7. This fingerprints consistently correctly, at least for me.
"And how effective do you think the new rules will be at preventing those younger than 18 from gaining access to pornography?"
-> 64% "not very effective / not at all effective"
Payment processors have major network effects in that infra setup is expensive, banks need to be onboarded one-by-one, and whichever network has the most consumers, businesses will gravitate towards it. Iterate this over 20 years, and this always results in natural monopolies / duopolies. This creates a natural chokepoint/linchpin over which millions of people's mutually exclusive needs are getting banged at; including consumers at large, govs at large, and special-interest groups at large.
Absent crystal clear legislation -and porn is anything, but- this will always be arbitrary, and leave one side in the dust.
OTOH: if the currently pending court case on anti-monopoly bars google from making payments to mozilla (which is about ~90%++ of their revenue), mozilla truly, and well is fucked. Meaning -they need to diversify, and they know it; they can't sell browsers, related services are heavily competed for, so ads & selling user data is broadly the only viable strat that can underwrite their existence.
Of course, the community won't have it. And therein lies the rub: by going with google's bribe, on this long term, they wrote themselves into a corner they can't exit.
O1 is higher quality, more nuanced, and has deeper understanding; the biggest downside rn is the significantly higher latency (both due to thinking, and also, continue.dev doesn't support o1 streaming currently, so you're waiting until it's all done), and higher cost.
In terms of tools: either vscode with continue.dev / cline, or cursor
Languages: node.js / javascript, and lately c# / .net / unity
There are many, many people, and companies who operate under the false belief that the CAN-SPAM act does not apply to them; and eg create new mailing lists to blast many people with their spam. Some of these unfortunately includes corps I have business relationship with (looking at you, Google), so "mark as spam" doesn't work well. Cease and desisting their legal department does. I have changed marketing strat of multiple largecorps by being a dangerous professional.
Once it's starts happening, speak; if speaking doesn't work, fight; if fighting doesn't work, move. This works.
* Given an annually compounding 30% linkrot, 99.92% of all the content ever published on the Internet is no longer available.
* This has been litigated, see Field v. Google Inc., 412 F (2006), and held to be "fair use" due to safe harbor of Section 512(b) of the DMCA
* This exemption does not apply to books, music, videos, or any of the other pirated material.
For larger context, the ecosystem is fragmenting, and I have ~10 browser extensions that are critical to me. I don't think I will prioritize chrome's software cadence over my own preferences, thank you.
For larger context, the ecosystem is fragmenting, and I have ~10 browser extensions that are critical to me. I don't think I will prioritize chrome's software cadence over my own preferences, thank you.
https://chatgpt.com/share/d5709aeb-d24c-488b-985c-c13eba0c01...
"4. IORP Directive: The IORP (Institutions for Occupational Retirement Provision) Directive is analyzed, highlighting its scope and its impact on pension funds across the EU. The paper suggests that the directive's complex regulations create inconsistencies and may need clarification or adjustment to better align with national policies." "5. Regulatory Framework and Proposals: A significant portion of the paper is devoted to discussing potential reforms to the regulatory framework governing pensions in the EU. It proposes a dual approach: a "soft law" code for non-economic pension services and a "hard law" legislative framework for economic activities. This proposal aims to clarify and streamline EU and national regulations on pensions."
^^ these corresponds to the author's self-selected two main points.
Asking because in the vast majority of cases, the phishing landing page has way more signals to recognize than the email headers.
What inspired that question?
My counterparties are not real.
In descending order of frequency:
* Tirekickers / wannabes: these people have fantasies about doing a startup... someday, but definitely not just now.
* Super excited about <thing/area>... has no related experience, fails at basic business ontology ("target market", "valueproposition") <- 95% mark
* Has no hypothesis about marketing channels, nor any insight on why this particular combination might work
* fails on all of market scoping, TAM, customer development, financial model <- 99% mark
Rest: limited operational experience, OR self-defeating / low psychological resilience, going nowhere.
There is a laundry list on sibling comment (https://news.ycombinator.com/item?id=39904704) for ticking boxes. My current hypo, is that peeps who check these boxes AND don't have a tech cofounder on their rolodex typically go to angels/VCs, and get a recommendation from them; and therefore will never appear on any markets for cofounders. Curious if this matches your experience.