> The communication must be made in a manner in which another person may view it.
Even 'transmitted' is too broad if you also consider iCloud backup to be a means.
9,289 karma · joined April 23, 2012
> The communication must be made in a manner in which another person may view it.
Even 'transmitted' is too broad if you also consider iCloud backup to be a means.
The providers[1] behind this Web Search API have very different rates:
Ceramic.ai: $0.25 per 1,000 requests
Linkup: $5.00 per 1,000 requests
Exa: $7.00 per 1,000 requests
[0] https://serper.dev/> Gadgets you build with the Muse Gadget SDK need a token. Add it to your SDK configuration so your gadgets can pair with the Muse app.
> Each token can be used by a limited number of devices.
> The Muse Gadget SDK is intended for personal tinkering and is not a supported product or developer platform. The SDK can change, break, or stop functioning without warning.
Doesn't really feel like it's my own.
This in itself I find problematic. Solving the mundane and realistic safety challenges are what prevents the potential (preposterous) outcomes. There's no need to ignore anything other than the sensationalism like anthropomorphising. The Paperclip Maximizer is meant to be an exaggerated thought experiment where an AI doesn't actually "want" anything, it's "literally doing" what it's supposed to. It's surprising that Pinker doesn't seem to see that, or sees it and doesn't believe no version of it is possible.
Construction machinery doesn't need to be alive to kill you--just left running unattended.
The difference is that AI doesn't have a well bounded construction site.
These numbers could and should get much better. As an example I can run Qwen3.8-27B-MXFP4 (W4A8) on 2x AMD R9700 that gets 260+ tokens/sec to start and slows down to ~110 tokens/sec over 128k context and can do the max 256k. These are for batch size 1 and throughput goes higher with batching. This is due to speculative decoding, efficient all-reduce inter-gpu compression, and custom GEMM kernels for the specific hardware. Note each R9700 only has 644 GB/s memory bandwidth.
Similar for LLM measures from an ideal 1.0 mark.
Apparently they're cooking up MAI (Microsoft AI), and I'd seen their small Phi models listed online. Calling Anthropic a competitor is hilarious.
I also believe there are limitations of LLMs, but not necessarily where people think. I won't expect LLMs to be creative solvers until they can tell a novel funny joke with any recognition/consistency.
Another good explanation here is that it's all taking place but that last step of creating the visual is suppressed (as intended) while conscious as that would be hallucinating.
- 252 GB HBM3e VRAM
- 496 GB LPDDR5X RAM
I would rather have a system using 2x Instinct MI350P GPUs (288GB total) for much less.At work mostly Opus 4.8 (sometimes a GPT or Gemini 3.1 Pro). I find Opus 5 chatty/slower and Fable can venture into over-engineering itself into unnecessary complications.
> Languages: TypeScript 98.4%, Other 1.6%
Not for me.
> What is the internet now to me? In many ways it’s something that I have to tolerate to do many things that functioned fine before. I have to use an app to pay for parking. I need to create an online account to pay for a family swim at the local leisure centre. I have to submit an online form to book a slot at the local recycling centre. I need an app to check my bank and credit card balances.
It's so much more annoying to do any of these things without/before internet. I don't do half of the other things in the rest of that paragraph. Like for phones I just get a used CAD$400 Android phone with good audio output.
I just got DeepSeek Harness (DSH) set up with 2x R9700 and it's rather mind blowing that these can do actual work and quickly. Up until now I've always been evaluating and searching for better hardware/model/tweaks. This is much more than I even hoped for and considered getting extra 3090/4090. Now I can stop looking/tweaking and start using it for all the different things I've yet to discover it's good for. I do plan to also try/use Hermes and Pi. DSH is annoying that every plugin install/remove requires a restart--given that "everything's a plugin".
What kind of performance are you getting with 4x R9700s--what do you do with all the VRAM (batching, concurrent requests, etc)?
That was definitely the case of the Unity desktop that really only worked well on netbooks. That's when I lost confidence in Ubuntu for design. Loss in Canonical on the whole came later.
Each MI350P[1] in the TR Halo Station has 144GB VRAM and with 4.6 PFLOPs peak MXFP6 performance.
Four liquid cooled? Yes please.
[0] https://a16z.com/building-a16zs-personal-ai-workstation-with...
[1] https://www.amd.com/en/products/accelerators/instinct/mi350/...