1,035 karma · joined September 6, 2012
Self hosting, SBCs, AI/Vision/LLMs
Once you referenced routers, middleware and services simply by name, but that changed into per-source scoped versions (e.g. service1@file, middleware@docker).
I kept bumping into those edge cases (custom SSL cert set-up was really confusing), but thanks to chatgpt, I at least ended up with workable solutions.
Plus, making ad blocking a channel owner's problem is kind of genius.
It's nice for simple stuff, but wouldn't recommend for anything needing precision. In the beginning it was laying out perfect layers, but after 6 months just lost it. Been doing every possible maintenance stuff and it just doesn't go back.
When I started Uni, the "A diploma will guarantee you great job opportunities" mantra was unshakeable.
Now I think, the pendulum swung so hard in the other direction, that kids of same age have tons of refuttals at their disposal. It must take a lot more work, from parents, to instill and motivate what was once seen as a good career starter.
Humans do it too. I have given up on my country's non-local information sources, because I could recognize original sources that are being deliberately omitted. There's a satiric webpage that is basically a reddit scrape. Most of users don't notice and those who do, don't seem to care.
When I started using it (~ 2 years) , it was necessary. Google was simply not solving any of my actual issues (software related).
Now, It seems that google might have improved a bit. I check from time to time and the gap isn't as huge, as when Kagi started
I assume we can go up to 120B using fp8?
It's amazing how much can be done, but anything that I think will take an hour, lasts 6. Many times into the night.
"Sure, I'll help you stop flirting with OOMs"
"Thought for 27s Yep-..." (this comes out a lot)
"If you still graze OOM at load"
"how far you can push --max-model-len without more OOM drama"
- all this in a prolonged discussion about CUDA and various llm runners. I've added special user instructions to avoid flowery language, but it gets ignored.
EDIT: it also dragged conversation for hours. I ended up going with latest docs and finally, all issues with CUDA in a joint tabbyApi and exllamav2 project cleared up. It just couldn't find a solution and kept proposing, whatever people wrote in similar issues. It's reasoning capabilities are in my eyes greatly exaggarated.
Was able to create a sample page, tried starting a server, recognising a leftover server was running, killing it (and forced a prompt for my permission), retrying and finding out it's ip for me to open in the browser.
This isn't a demo anymore. That's actually very useful help for interns/juniors already.
I had gpt-5 only on my account for the most of today, but now I'm back at previous choices (including my preferred o3).
Had gpt-5 been pulled? Or, was it only a preview?
YouTube on the other hand...
That's 3 medications?
Also, how convenient those stories come out in light of upcoming "regulatory" safekeeping measures.
This whole article reads like 4-chan greentext or some teenage fanfiction.
That looks like a version number...
Would like to see more of the captured data, because a simple "about" dialog, would also need to call some server to check, if it software is in the latest version. To display the "you have the latest version" label.
Hello,
Your Pro plan just got way more powerful with three major upgrades previously available only to Max, Team, and Enterprise users.
Claude Code is now included
Claude Code is a command line tool that gives you direct access to Claude in your >terminal, letting you delegate complex coding tasks while maintaining full control. You can now use Claude Code at no extra cost with your Pro subscription.