HNHacker News
TopNewBestAskShowJobs

mikeravkine

164 karma · joined May 27, 2023

submissionscomments
mikeravkine··on LLM Leaderboard with explanations of what each score means
The Internet is 90% mobile users that can't hover.
mikeravkine··on Towards 1-bit Machine Learning Models
Thank you for your efforts on behalf of the GPU poor!

It's getting tougher to use older, cheaper GPUs (Pascal/Maxwell) with modern quantization schemes so anything you can do to keep kernels compatible with SM52 and SM61 would be greatly appreciated.

mikeravkine··on VSCode Drops Ubuntu 18.04 Support
Ubuntu 18 went EOL on May 31, 2023 so about 8 months ago. What CentOS are you running that's still in support and was affected by this change? I had to upgrade all my systems last year after the EOL as most python packages dropped 3.8 support.
mikeravkine··on Why isn't Bluesky a peer-to-peer network?
The challenge to Mastodon isn't adoption, its absense is a symptom of a deeper issue: it's a Very Bad Idea to trust random server operators with both your data and it's sovereignty.

You can take your data to another server only if the original server is up. When the admin forgets to pay the bill and ISP nukes the box and the backups admin said they had don't exist and the coadmin turns out to have no access at all, you lose everything. Ask me how I know.

Self hosting in theory should fix this but it doesn't: it's a resource intensive afterthought because the developers only really care about their Big Instance.

mikeravkine··on Show HN: Voxos.ai – An Open-Source Desktop Voice Assistant
I just wanted to say thank you, aider (with the new unified diff format) is the first AI tool that has actually changed the way I work.
mikeravkine··on Canadian man stuck in triangle of e-commerce fraud
The razor doesn't apply to the police.
mikeravkine··on Ditching PaaS: Why I Went Back to Self-Hosting
Hetzner can and will terminate your account and delete all your data on a whim, never telling you why. There is nothing at all you can do to prove your usage was legitimate. I get that they must deal with a lot of abuse but I wouldn't use them for anything mission critical ever.
mikeravkine··on Analysis of 200M newspaper pages: Sentiment has collapsed over the past 50 years
In terms of economics specifically, in what way have things "gotten better"? Purchasing power has collapsed, the poor sentiment here seems the correct sentiment to me.
mikeravkine··on Using Alpine can make Python Docker builds 50× slower
There is no difference in download time for 20mb vs 50mb in a datacenter, but the amount of time lost to alpine being "quirky" is incalculable.
mikeravkine··on Running Llama.cpp on AWS Instances
If anyone is looking for a more reasonably cost effective solution, Hetzner has 16 vCPU/32GB RAM ARM VMs for $24E/mo that will run 34b Q4 GGUF at around 4 tok/sec. It's not very fast, but it is very cheap.
mikeravkine··on Get started with technical writing
Where do I get one of these tech writing destroying AIs that will both understand how to use my product and produce clear, concise guides for doing so without burdening the existing team?
mikeravkine··on Please ignore the deluge of complete nonsense about Q*
This is not about intelligence, it's about agency. Text completion generators don't have agency no matter how big they get.
mikeravkine··on Fast Llama 2 on CPUs with Sparse Fine-Tuning and DeepSparse
Hetzner offers incredibly cheap ARM machines in the Falkenstein DC, for 25Eur a month you can snag the top of the line with 16 vCPU and 32GB RAM.

If your usecase fits inside that 32GB (no 70B models, sadly) the price to performance of a GGUF Q4KM is really attractive on this setup.

mikeravkine··on Show HN: New visual language for teaching kids to code
Pascals := is concise and intuitively obvious, so obviously it never caught on.
mikeravkine··on 'I'm a Doomer': OpenAI's New Interim CEO Wants to Slow AI Progress Down
1997 era internet would have scared the shit out of you.
mikeravkine··on Ask HN: Why is OpenAI firing Sam Altman such a big deal?
Why would you need to compete with such an entity? What does "compete" even mean here? Can't instead everyone get one (or 10) of these to.. assist them? I don't understand why we are jumping directly into dystopian scifi vs what past tool revolutions actually look like.
mikeravkine··on Ask HN: How do you personally measure the success of your open source projects?
Receiving a PR is a massive success in my books, it means someone not only found it useful but also improved it! The win-win of open source.
mikeravkine··on Ask HN: Is anyone using cloud dev environments (e.g. Codespaces/Replit) at work?
With VS Code remote SSH, there is no "local" you are always on the server so there is also no syncing. They do some tricks to make this seamless and perform and feel as if everything was local.
mikeravkine··on I’m not a programmer, and I used AI to build my first bot
I always write the fun bits of code myself, it's usually easier then explaining to the model what I want anyway.. but like 80% of any project is scaffolding or plumbing, The Machine can do that for me with 10x speed and precision without taking any enjoyment away from me.
mikeravkine··on Show HN: Carton – Run any ML model from any programming language
..isn't this just Docker?
mikeravkine··on X scraps tool to report electoral fake news
It's now called Xitter (with the X pronounced 'sh')
mikeravkine··on Ask HN: 6 months later. How is Bard doing?
Canadian here, just checked and still no access. It's like some kind of bad joke is being played on us, Google can pound sand.
mikeravkine··on Signals says cloud repatriation has already saved $1M
Hetzner has some weird and very strong opinions on what you can and cannot do on their servers, likely because they are so cheap they attract all sorts of shady customers.

I've had good luck with OneProvider.

mikeravkine··on Show HN: finetune LLMs via the Finetuning Hub
OobaBooga supports this kind of load-and-go LORA: https://github.com/oobabooga/text-generation-webui
mikeravkine··on Youtube2Webpage: Create Websites with Text from Videos
I've been working on just such a tool [1] to help me digest podcasts and senate hearings.

[1] https://github.com/the-crypt-keeper/tldw

mikeravkine··on Phind (code beating GPT4) seems to have used WizardLM's finetuned checkpoint
Mismatched parenthesis, specifically using )) instead of ) is a common failure mode I observed in all of my codellama testing.
mikeravkine··on Whisper.api: Open-source, self-hosted speech-to-text with fast transcription
One caveat here is that whisper.cpp does not offer any CUDA support at all, acceleration is only available for Apple Silicon.

If you have Nvidia hardware the ctranslate2 based faster-whisper is very very fast: https://github.com/guillaumekln/faster-whisper

mikeravkine··on How Is LLaMa.cpp Possible?
I have several sets of quant comparisons posted on my HF spaces, the caveat is my prompts are all "English to code": https://huggingface.co/spaces/mike-ravkine/can-ai-code-compa...

The dropdown at the top selects which comparison: Falcon compares GGML, Vicuna compares bits and bytes. I have some more comparisons planned, feel free to open an issue if you'd like to see something specific: https://github.com/the-crypt-keeper/can-ai-code

mikeravkine··on Fine-Tuning Llama-2: A Comprehensive Case Study for Tailoring Custom Models
The model card also has prompt formats for context aware document Q/A and multi-CoT, using those correctly improves performance at such tasks significantly.
mikeravkine··on Blocked by Cloudflare
I have noticed that on StarLink some sites behind CF go into "prove you are human" loops that are impassable.

What causes such loops? Just a challenge over and over.

Page 1 of 2Next →