HNHacker News
TopNewBestAskShowJobs

electroglyph

743 karma · joined September 29, 2018

admin@terminoid.com
submissionscomments
electroglyph··on AI is code – and can't be prompted into being smarter
it's probably a lie
electroglyph··on DeepSeek V4 Pro beats GPT-5.5 Pro on precision
deepseek 4 pro is insanely good for the price
electroglyph··on Nvidia is proposing a beast of a CPU system for Windows PCs
that link actually recommends not doing it from UEFI and doing it via software
electroglyph··on KVarN: Native vLLM backend for KV-cache quantization by Huawei
any divergence (even if the benchmark is better) from full precision is error
electroglyph··on Gustav Klimt and Egon Schiele in Conversation (2018)
this is better than TFA
electroglyph··on FBI Arrests CIA Official with $40M in Gold Bars in His Home
sometimes i wonder if the left hand knows what the right is doing. it looks like we arrested our own spy in this case: https://www.politico.com/news/2026/05/25/american-journalist...
electroglyph··on Norway's 2 petabytes of Huawei flash storage and LLM training
absolutely. somebody online was wanting an LLM with Georgian language support, and that's exactly what i suggested: start digitizing Georgian text.
electroglyph··on Wake up! 16b
i'll upvote this each time it's submitted
electroglyph··on Qwen3.7-Max: The Agent Frontier
you should be using dflash with that model, look it up
electroglyph··on DeepSeek-V4-Flash means LLM steering is interesting again
heretic maintainer: https://github.com/p-e-w/heretic

the fun bits are in another branch or PRs

electroglyph··on DeepSeek-V4-Flash means LLM steering is interesting again
p-e-w was just talking about this the other day in his Discord. seems doing the one neuron method is quite bad for KLD and that's why the newer techniques have stuck.
electroglyph··on Seeing Birdsong
site has so little information there doesn't seem to be much to discuss
electroglyph··on A polynomial autoencoder beats PCA on transformer embeddings
this looks awesome. i've been struggling with vector compression, and have been trying PCA + all sorts of rotations. looking forward to trying this out
electroglyph··on Making LLM Training Faster with Unsloth and NVIDIA
nice writeup! looking forward to doing some more training as soon as i get some more data sorted. it'll be a custom arch, but i'll probably shoehorn it into unsloth for a speed boost.
electroglyph··on Train Your Own LLM from Scratch
you can train it, but not fully
electroglyph··on Opus 4.7 knows the real Kelsey
that's in the ideal scenario where it's only seen a single copy of it tho
electroglyph··on New copy of earliest poem in English, written 1,3k years ago, discovered in Rome
it was 1.3e-6 billion years ago!
electroglyph··on Microsoft and OpenAI end their exclusive and revenue-sharing deal
i'm doing inference on a free mi300x instance from AMD right now. not sure if the software stack is just old or what, but here's what i've observed: stuck on an old version of vllm pre-Transformers 5 support. it lacks MoE support for qwen3 models. oss-120b is faaaar slower than it should be.

int8 quantization seems like it's almost supported, but not quite. speeds drop to a fraction of full precision speed and the server seems like it intermittently hangs. int4 quantization not supported. fp8 quantization not supported.

again, maybe AMD is just being lazy with what they've provided, but it's not a great look.

right now the fastest smart model i can run is full precision qwen3-32b. with 120 parallel requests (short context) i'm getting PP @ 4500 tokens/sec and TG @ 1300 tokens/sec

electroglyph··on Agents Aren't Coworkers, Embed Them in Your Software
but should you drive or walk to the car wash?
electroglyph··on Lambda Calculus Benchmark for AI
i dunno, Opus is losing it's edge imo. i regularly use a mix of models, including Opus, glm 5.1, kimi 2.6, etc. and i find that all of them are pretty much equally good at "average" coding, but on difficult stuff they're nearly equally bad. i can't deny that Opus has an edge, but it's not a huge one.
electroglyph··on Which one is more important: more parameters or more computation? (2021)
> they also don't know what they don't know

they sort of do tho:

https://transformer-circuits.pub/2025/introspection/index.ht...

electroglyph··on Plain text has been around for decades and it’s here to stay
how about a unicode art tool?

https://electroglyph.github.io/atheriz_draw/

electroglyph··on OpenClaw isn't fooling me. I remember MS-DOS
https://sleepingrobots.com/dreams/stop-using-ollama/
electroglyph··on I ran Gemma 4 as a local model in Codex CLI
flow matching is making some strides right now, too
electroglyph··on Sam Altman's response to Molotov cocktail incident
i don't buy this. distilled how? you don't get access to logprobs, and the thinking traces are fake and compressed. it's an expensive way to get potentially substandard training data.
electroglyph··on Creating the Futurescape for the Fifth Element (2019)
nah, a crypto grifter released one with cooked benchmarks
electroglyph··on GLM-5.1: Towards Long-Horizon Tasks
better than Opus? not even close. after struggling thru server overload for the past couple hours i finally put 5.1 thru the paces and it's....okay. failed some simple stuff that Sonnet/Opus/Gemini didn't. failed it badly and repeatedly actually. this was in typescript, btw. not sure if i'll keep the subscription or not
electroglyph··on Caveman: Why use many token when few token do trick
after you go from from millions of params to billions+ models start to get weird (depending on training) just look at any number of interpretability research papers. Anthropic has some good ones.
electroglyph··on $500 GPU outperforms Claude Sonnet on coding benchmarks
what's with the weird "Geometric Lens routing" ?? sounds like a made up GPTism
electroglyph··on Autoresearch on an old research idea
ah, you've found the danger zone!
← PreviousPage 2 of 6Next →