HNHacker News
TopNewBestAskShowJobs

sabareesh

323 karma · joined April 12, 2017

CTO @ guidedchoice.com , 3nickels.com . Making finance easy for everyone
submissionscomments
sabareesh··on 45°C cooling design cuts data center water use to near zero
I am pretty much doing the same but running the coolant at 40 deg C instead of 45 as my pumps are rated for 45 C max temp. Here is bit more about my setup https://sabareesh.com/posts/blackwell-waterblock/
sabareesh··on 4× RTX Pro 6000 Blackwell on Water, and the One Card That Wouldn't Behave
Yes this has been on my mind as well. But this was built one at a time but still overall very happy with them
sabareesh··on 4× RTX Pro 6000 Blackwell on Water, and the One Card That Wouldn't Behave
Appreciate the feedback. I have improved the article.
sabareesh··on 4× RTX Pro 6000 Blackwell on Water, and the One Card That Wouldn't Behave
Appreciate the feedback. I have improved the article.
sabareesh··on 4× RTX Pro 6000 Blackwell on Water, and the One Card That Wouldn't Behave
I am primarily experimenting on post training stack. As of now working on training a model that is natively RLM https://github.com/alexzhang13/rlm
sabareesh··on 4× RTX Pro 6000 Blackwell on Water, and the One Card That Wouldn't Behave
It is basically on 2 different circuits/breakers. Asus wrx90e supports 2 psu as well. You may need to synchronize both psu and several adapter for this is available in Amazon. Soon planning to upgrade it to 240V
sabareesh··on 4× RTX Pro 6000 Blackwell on Water, and the One Card That Wouldn't Behave
Not sure what really happened but some force or bad solder caused it.
sabareesh··on 4× RTX Pro 6000 Blackwell on Water, and the One Card That Wouldn't Behave
Most of the training i am working on is with post training. You can do so much with a system that is running 24/7
sabareesh··on 4× RTX Pro 6000 Blackwell on Water, and the One Card That Wouldn't Behave
Sure 140mm fans you may call little but it does need enough static pressure for the radiators. This setup is already several times quieter than stock setup
sabareesh··on 4× RTX Pro 6000 Blackwell on Water, and the One Card That Wouldn't Behave
Converting four RTX PRO 6000 Blackwell cards to waterblocks, finding a VRM choke loose on the workbench, and getting back to 41k tok/s.
sabareesh··on KV Cache Compression 900000x Beyond TurboQuant and Per-Vector Shannon Limit
Sounds like speculative decoding but for KV cache
sabareesh··on Claude Opus 4.7
Based on last few attemts on claude code to address a docker build issue this feels like a downgrade
sabareesh··on OpenAI's latest repo has Claude as the third top contributor
Codex usually dont add itself as contributor so this is misleading .
sabareesh··on Nvidia greenboost: transparently extend GPU VRAM using system RAM/NVMe
I wish it provided benchmark comparing Direct RAM offload vs CPU offload vs Full VRAM
sabareesh··on Self Driving Car Insurance
Tesla have their own Insurance product which is already very competitive compared to other providers. Not sure if lemonade can beat them . Tesla's insurance product has similar objective in place already where it rewards self driving over manual driving.
sabareesh··on Stop Doom Scrolling, Start Doom Coding: Build via the terminal from your phone
I am looking for some open source terminal for iphone .I have code server running which i can just use terminal from vs code on safari
sabareesh··on I switched from VSCode to Zed
Sorry to disappoint. But purely codex and claude code
sabareesh··on I switched from VSCode to Zed
I have switched to terminal
sabareesh··on IQuest-Coder: A new open-source code model beats Claude Sonnet 4.5 and GPT 5.1 [pdf]
TL;DR is that they didn't clean the repo (.git/ folder), model just reward hacked its way to look up future commits with fixes. Credit goes to everyone in this thread for solving this: https://xcancel.com/xeophon/status/2006969664346501589

(given that IQuestLab published their SWE-Bench Verified trajectory data, I want to be charitable and assume genuine oversight rather than "benchmaxxing", probably an easy to miss thing if you are new to benchmarking)

https://www.reddit.com/r/LocalLLaMA/comments/1q1ura1/iquestl...

sabareesh··on Show HN: Stop Claude Code from forgetting everything
Non starter for us, we cant ship propriety data to a third party servers.
sabareesh··on Gemini 3 Flash: Frontier intelligence built for speed
this has one of the worse score in AA-Omniscience Hallucination Rate
sabareesh··on Gemini 3 Flash: Frontier intelligence built for speed
Nope lower is better compared to recent open ai models this is bad. I am looking at AA-Omniscience Hallucination Rate
sabareesh··on Gemini 3 Flash: Frontier intelligence built for speed
Watch out these model are hallucinating lot more https://artificialanalysis.ai/evaluations/omniscience?omnisc...
sabareesh··on The Big Vitamin D Mistake [pdf] (2017)
So is 10,000 IU of daily does ok ?
sabareesh··on When Tesla's FSD works well, it gets credit. When it doesn't, you get blamed
Technically you kind of get this in Nevada when using Tesla insurance and if you drive 100 % FSD. If you drive manually you are pretty much doxed for random Front collision Warning which is super sensitive
sabareesh··on Karpathy on DeepSeek-OCR paper: Are pixels better inputs to LLMs than text?
It might be that our current tokenization is inefficient compared to how well image pipeline does. Language already does lot of compression but there might be even better way to represent it in latent space
sabareesh··on Claude Code on the web
Similar feeling. Seems it is good at certain things and if something doesnt work it want to do things simply and in turn becomes something that you didnt ask for and certain times opposite of what you wanted. On the other hand with codex certain time you feel the AGI but that is like 2 out of 10 sessions. This is primarily may be due to how complete the prompt and how well you define the problems.
sabareesh··on The Tiny Teams Playbook
"Simple, Boring Tech Stack:" Good advice but bad example, because it depends on what engineers are familiar and comfortable with and technology itself should be mature enough. You dont want to spend time building orchestrator when k8s solves it for you. Most cloud provides provide you with k8s as a service, which are miles better than using shell scripts, if you are already familiar with k8s
sabareesh··on Llama-Factory: Unified, Efficient Fine-Tuning for 100 Open LLMs
This is great,but most work is involved in curating the dataset and the objective functions for RL.
sabareesh··on Nvidia's new 'robot brain' goes on sale for $3,499
Looks very similar to DGX spark
Page 1 of 6Next →