HNHacker News
TopNewBestAskShowJobs

mudkipdev

1,039 karma · joined August 9, 2023

submissionscomments
mudkipdev··on Qwen3.6-27B: Flagship-Level Coding in a 27B Dense Model
HuggingFace has a nice UI where you can save your specs to your account and it will display a checkmark/red X next to every unsloth quantization to estimate if it will fit.
mudkipdev··on Qwen3.6-27B: Flagship-Level Coding in a 27B Dense Model
Why is the assumption that they trained for a pelican on a bicycle, rather than running RL for all kinds of 'generate an SVG' tasks?
mudkipdev··on Claude Code to be removed from Anthropic's Pro plan?
The GLM coding plan price increased dramatically
mudkipdev··on A Roblox cheat and one AI tool brought down Vercel's platform
I'm getting a "failed to verify your browser" error on this article
mudkipdev··on Claude Token Counter, now with model comparisons
Why do you need an API key to tokenize the text? Isn't it supposed to be a cheap step that everything else in the model relies on?
mudkipdev··on Six Levels of Dark Mode (2024)
Grayish dark themes are underrated
mudkipdev··on Changes in the system prompt between Claude Opus 4.6 and 4.7
The Claude prompt is already quite bloated, around 7,000 tokens excluding tools.
mudkipdev··on State of Kdenlive
If anyone has a better workflow for creating lots of captions in kdenlive please let me know. I had to duplicate each title to the media library and drag it into the timeline, because if I simply copy/pasted then the text content/styling would be shared across instances
mudkipdev··on Qwen3.6-35B-A3B: Agentic Coding Power, Now Open to All
Re-read that
mudkipdev··on I ran Gemma 4 as a local model in Codex CLI
Does the large system prompt work fine for this model? If needed, you could use a lightweight CLI like Pi, which only comes with 4 tools by default
mudkipdev··on Ask HN: What Are You Working On? (April 2026)
I built a Claude-inspired UI for Ollama/llama.cpp

https://github.com/mudkipdev/chat

mudkipdev··on Show HN: I built a tiny LLM to demystify how language models work
People have made toki pona translation models before, not exclusively trained though
mudkipdev··on Gemma 4 on iPhone
It's strange that my iPhone 14 is at regular temperature when using the E2B model. But also it's a lot slower (not sure how to measure the exact tokens per second, ~12 if I had to guess)
mudkipdev··on Show HN: I built a tiny LLM to demystify how language models work
This is probably a consequence of the training data being fully lowercase:

You> hello Guppy> hi. did you bring micro pellets.

You> HELLO Guppy> i don't know what it means but it's mine.

mudkipdev··on Google releases Gemma 4 open models
If you use the 'run' command, it pulls automatically for you
mudkipdev··on Google releases Gemma 4 open models
Under 15 is too slow for conversation personally. I guess 5 tokens per second is nice if you're one of the people who likes letting coding agents run overnight
mudkipdev··on Google releases Gemma 4 open models
Can't wait for gemma4-31b-it-claude-opus-4-6-distilled-q4-k-m on huggingface tomorrow
mudkipdev··on r/programming bans all discussion of LLM programming
There can't be any interesting discussion about AI programming. Every conversation boils down to what skill files you use, or how Opus 4.6 compares to Codex, or how well you can manage 16 parallel agents.
mudkipdev··on Gemini 3.1 Flash Live: Making audio AI more natural and reliable
Not to be confused with Gemini 3.1 Flash Lite
mudkipdev··on Chat GPT 5.2 cannot explain the German word "geschniegelt"
It simply means the tokenizer's training corpus may have included a massive amount of German literature or accidentally oversampled a web page where that word was frequently repeated. Look up "glitch tokens" to learn more.
mudkipdev··on Is the Future of AI Local?
Yes, I agree that the value of Apple's chips is hard to beat, but there's still a massive bottleneck with hardware accessibility. The 128 GB MacBook described costs over $5,000 on their website, and in the consumer space, even the most often recommended GPU which is a 3090 with 24 GB VRAM, you can find used going for $700 at minimum. This effectively prices out all non-professional users who don't have money to spend to simply have parity with the $20 subscription on their phones. (and even for those who do, they have to cope with the fact that the model will always be dumber than ChatGPT, and that their hardware will grow outdated very quickly)

Also a noticeable disconnect between the hardware we have and the primary focus of open-source labs (scale up and cater to their enterprise customers, just look at GLM-5's increase to 744B parameters, double from GLM-4.5's 355B). We really just need some kind of Cambrian explosion in cheap hardware for local models to be feasible.

mudkipdev··on What 81,000 people want from AI
This page without exaggeration reduced my browser to 5 frames per second.
mudkipdev··on GPT‑5.4 Mini and Nano
This is a bot
mudkipdev··on Ask HN: How is AI-assisted coding going for you professionally?
Are you saying you're learning go because you've freed up time elsewhere or is AI helping?
mudkipdev··on A most elegant TCP hole punching algorithm
This is an AI slop bot
mudkipdev··on Writing my own text editor, and daily-driving it
I would recommend using the ropey crate for easy performance gains. A string buffer is quick to implement but you will hit a wall as soon as you need to edit large files.
mudkipdev··on Ask HN: Please restrict new accounts from posting
I think vote rigging detection might be based on the length of your session
mudkipdev··on GPT-5.4
Probably refresh the api models list every couple minutes instead. No one could have guessed the name of GPT-Codex-Spark
mudkipdev··on A plastic made from milk that vanishes in 13 weeks
Or remove "is"
mudkipdev··on When does MCP make sense vs CLI?
This got renamed right in front of my eyes
← PreviousPage 2 of 5Next →