I don't think directly, but search as a business category was once a solved domain with a single winner, and suddenly it was up for grabs again with LLMs. If Google ever lost search, they would lose Ads which is their entire core business. Google 'won' in the sense that they successfully integrated AI into search fast enough and well enough that they will (likely) hold their place through this transition, and can still funnel ads to users which is already integrated into their AI overview.
AI has made me wonder a lot lately about what it is that actually creates conscious experience, what is the actual physical mechanism that produces 'experience' (or is there a single valid mechanism, or rather just some property that can be expressed many ways). The more I think about it, the more I realize I have no fucking clue, and the more interested in the question I get.
I've landed here as well, but by force. I've actually tried to share some of the productivity and workflow tools I've built with AI that help automate portions of my work, specifically on target firmware debugging, and everyone at work just...ignored it. Didnt care. Of the small number of people who even stopped to look, a few of them were actively negative. Maybe I didnt sell it hard enough, but my job isnt to sell productivity tools internally (or even make them, i did it for myself). So, I'll keep making myself more productive and build tools for myself. And everyone else will just have to figure out their own path. I tried to help, but oh well.
If he had one shortly after they launched, the suspension was truly terrible. R1T was clearly where all the work went and it's like no one even drove an R1S before they released it. Thankfully due to SW updates the suspension is absolutely night and day better, it went from making me violently sick to pretty nice. Assuming dudes car wasn't just a lemon, if he had and got rid of the R1S before the suspension got better I can 100% believe he hates the car.
I can understand when people say the code was the way part, but under the assumption that the system design and architecture are clean, code is high quality or the project is greenfield, and there is proper testing and validation. Then, sure, the lines of code aren't the hardest but that's only because that hardest work was front loaded and given a different name. Even then it's still not always easy.
Particularly if LLMs plateau and intelligence becomes essentially commoditized. If they can't compete on models alone they'll start moving more and more up the stack.
Exactly this. LLMs are already some form of useful, and have some kind/level of intelligence. We don't need to get to AGI before AI has uses. It isn't a step function and it isn't all or nothing. Even if LLMs get no better than they are today, we will still have valid practical applications that use them.
Also isn't the point moot if fable is so scary it needs to be banned and open weight models are already close on its heels? Even if China never imported another Nvidia chip the models he's so scared of are already out of the bag. At this point democratizing access seems like the best path forward.
Ive been using btrfs snapshots and some auto generated isolation rules plus a git ceiling at the mount root for the btrfs image (have to do this in wsl, stupid work computer). it's worked really well and fable hasn't had any issues with the "sandbox" (obviously not really but it works well enough)
I'd be curious to see the results, especially with some models having 1.5m and 2m context sizes, if the first 75% of the context was filled with unrelated info.
People don't seem to be able to reconcile the fact that there is likely an overbuild and overspend on AI that may be inflating a bubble, and that AI is actually incredibly useful and getting really really good for certain tasks. Both camps are right, except for when they say the other is wrong.
I think the intent was just to show how sensitive the classifier is. If it flags prompts that simple, there's no hope for anything biology related at all really.
And you can bypass a seatbelt warning by just plugging in a buckle without the belt, but most people don't bother. It's not worth the inconvenience to circumvent, so it still has a positive impact on safety.
This time is different though (which has also been said every single time). But I'm worried this time it's true (also said every time). Doesn't help with the unease though.
I actually agree, Linux is well past the point a minimally tech competent person can use it fine, but it doesn't solve the fact that even if Linux was flawless, there is still a switching cost in time, relearning a new system, and worst (best?) of all of you decide you are willing to do all of that now you can get entirely lost in the weeds picking a distro. I used Linux all through university, then went back to windows out of convenience and needing to use it for work anyways.
Until one day I got so frustrated with constant settings resets, reboots at the worst times for software updates that fail, highjacking my pc after every update for a guided tour of the latest things Microsoft decided to break, and telemetry that can only be disabled with an obscure registry hack that changes every few months, I just couldn't anymore.
Linux has been good enough as a daily driver for a while now, but even with proton I don't know if the pull factors towards Linux will ever be strong enough for most people. For me though the push factors away from Microsoft had gotten so strong I couldn't take it anymore.
Agreed. For personal use it's already easily worth $100 a month (to me personally). More probably. For work, it's entirely based on its financial impact for a given role, and for some people/companies it will be worth the cost even at $X thousand per month per seat.
A year ago this same guy was selling artificial super intelligence, right around the corner, and you'd get so rich if you would just give him some money. no idea why anyone believes the same guy when he pulls a new scam. I'm curious what he tries to pull next year.
I really like the 3d printer analogy. You can still make some pretty cool stuff, you can make some pretty complex things if you carefully design the whole system and put in the effort to print each part individually, and quality depends both on how good of a 3d printer you have and on the proper use of it. "3d print me a new house" is still a pipe-dream: you'll get some miniature facsimile of a house, sure, but a proper house requires proper tools and expertise.
Who cares? Nothing wrong with trying to make a product to sell, but projects dont have to be to sell. I've been having a blast lately working on an old game engine I started during covid and getting sidetracked into some new projects. None of them will ever make me a dime but I'm learning a ton and having fun.
I also think some of this stems from the default 1m context window. Performance starts to degrade when context size increases, and each token over (i think the level is) 400k counts more towards your usage limit. Defaulting to 1m context size, if people arent carefully managing context (which they shouldnt ever have to in an ideal world), they would notice somewhat degraded performance and increased token usage regardless.
Because most people work for someone else and don't decide their own salaries. It's not doubling productivity, but even a 10-20% boost to productivity for a team of engineers means that, as a business, even $1k per month per seat is perfectly acceptable. For consumers and hobbyists that basically kills access.
I think removing Claude Code from the $20 tier is a terrible idea, I never would've gone from nothing right into the $100/200 tier. The $20 plan let me get my feet wet and see how good it could be, and in less than a week I was on the $100 plan.
I think they need to at least have a 1 month introductory rate for the max plan at $20, or devs that decide to try out agentic coding just won't go to Anthropic.
That leads to downstream impacts, like when a company is deciding which AI coding tools to provide and the feedback management hears everyone is already used to (e.x.) Codex, then Anthropic starts losing the enterprise side of things.