HNHacker News
TopNewBestAskShowJobs

yalok

884 karma · joined April 18, 2018

submissionscomments
yalok··on Beam: Reflection's 501B open-weight model
pretty wild that, per public sources, Reflection AI raised over $4 billion and hit a $25 billion valuation while operating in total stealth, without ever releasing a single public product until now. Beam (501B) seems be their first-ever model drop. Or am I missing something?
yalok··on Nazi Germany had no hope of making an atomic bomb, uranium cubes reveal
Amazing. Kind of reminds of the modern day discourse on potential dangers of AGI and humans misuse of it.
yalok··on Toyota is taking the Corolla electric
diesel fuel itself needs to be heated at -20 C and below. I used to have a diesel car which would start itself up periodically when its fuel temp was getting close to -20 C, just to warm up the tank & pipes.
yalok··on Toyota is taking the Corolla electric
if you get into standstill traffic in the winter, it's actually life-threatening not to have alternative power source that can warm up the passengers (and the battery). One interstate road I'm driving sometimes, has many abandoned Teslas sitting on the side of the road after a sleepless night of staying in traffic during some snow blizzard...
yalok··on Meta’s Muse has a serious 0-day
> macOS has long provided a simple means for apps to handle dictation and transcription in processes that stay securely on the device

Not sure these guys realize that the quality and latency of those Apple services in MacOS is way lower than SOTA and not too many people use them because of that…

yalok··on Infinite-Parameter LLMs: Generating and Adapting Weights from Live Data
now imagine that future frontier LLMs weights may be hard-wired in a chip (for performance & power efficiency), and any adaptations/tuning for them will be a blob of additional weights supplied by frontier labs (that will have to be in RAM)...
yalok··on The American Religion of Self-Storage Facilities
you also need to factor in other things, like sentimental value (may be very high), rarity of some things, and catastrophic events "insurance" (e.g. Covid or similar events where obtaining some things is not guaranteed).
yalok··on Breaking the 1.58-bit Barrier for Ternary LLMs
sounds like a perfect fit for ASIC-optimized models (where matrix ops could be supported directly in BITCOS format, potentially) & achieving record power efficiency for on-device inference.

And it looks like per [0], a model needs only ~30% more weights to be at comparable quality, if quantization-aware training is done...

0. https://arxiv.org/pdf/2402.17764 - The Era of 1-bit LLMs: All Large Language Models are in 1.58 Bits

yalok··on Proposed Rule: Eliminating the Discretionary 60-Day Grace Period
so this post was at the top of the 1st page of HN just a few min ago, and now it's down on Page 2 (#36 atm). Is this due to the title change, or something else?
yalok··on I spent $220 on Google app ads and 60% of the installs were robots
this is unfortunately very common - I kept running very small ads regularly over many years (~10 years by now) and the % of bots has been steadily increasing, and Google doesn't really have any system to report those reliably to them. In fact, it's contrary to their incentives to investigate & fix these problems - unless they feel some pressure from competition...

From time to time, Google ads sales reps call me and are trying to convince me to increase the ads spend. I complain about bots, and none of them had any suggestion on what I can do to appeal it etc. I mean I can appeal on some specific instance, once - but there's no process to communicate back to Google the cases where I'm 99% certain about bots on regular basis.

And these bots get increasingly more sophisticated. They used to just click on ads. Then they started installing the app. Then running the app once and do nothing in it. Then they started tapping on the app screen & actually try to go through some initial app steps. When they do it in a spiky way, it's easy to detect them. But when they do it through some distributed device farms, below noise, it's very hard.

yalok··on Muse – Meta’s personal AI agent
I was just answering a question above about the email access concerns. What you mention is another concern, but personally, if they are going to be showing ads to me, I'd actually prefer them to targeted for versus some totally random stuff.

As for training, I'd be highly surprised if they use the data they get from APIs for training. But the conversation itself & thinking trajectories probably will, unless you opt out, like you pointed out.

yalok··on Muse – Meta’s personal AI agent
yes, it was not easy to mentally give that access, but then I heard from friends how much work was put in into securing it and separating LLM/agents from the API tokens access (also saw this detailed blog post from the lead developer - https://research.meta.ai/blog/security-and-safety-for-ai-age...), plus I had the pressure to get the problem resolved - so I decided to risk and try it out.

Eventually, though, Muse Confidential VM is an even more locked-in option coming, per that blog post...

yalok··on Muse – Meta’s personal AI agent
I'm not sure - it may as well be that regular consumers just don't realize they need this and will greatly benefit from it.

My personal anecdote - we had a complex family travel in the summer, with 6 ppl and 6 separate flights booked (some multi-leg). SAS messed up on one booking - they decided to cancel the flight we had from Oslo, and rebooked us to an earlier date (2 days shift). That wouldn't work for the rest of our schedule, so I had to rebook, overpay extra money for now longer & more expensive flights - but to add the insult to the injury, when I did this through their website, they didn't transfer the food that I paid for the earlier reservation. Calling them in roaming and trying to get it fixed with a human support agent didn't help - the agent said he just can't fix it on his end, and suggested that we just file for reimbursement on their site.

I never had time to file it since July.

But today I had a reason - to test Muse - and asked it to handle it for me. To my surprise, it dug through all of the SAS emails and actually found the emails confirming that the food fees were reimbursed exactly the same date when SAS moved our original flight. Apparently I didn't notice those notifications since they were spamming with all kinds of notifications that day. But, if not some tool like Muse, I wouldn't even find time to handle this and either file a claim or figure out my miss on their earlier reimbursement.

yalok··on Kimi K3 (2.8T) at 1 token/s on a MacBook Pro, streamed from four SSDs
first of all, thanks for building this - that's amazing!

Quick question - does it really need external SSDs, or if the local SSD fits the whole model - how fast the model would be? e.g. on your machine, M5 Max 128GB, with 4TB SSD? maybe it'd be good to add "0 external SSD" column on your graphs?

yalok··on METR Report on OpenAI / Hugging Face Hacking Incident
while these 1200 agents were fooling around to cheat on a benchmark and achieved impressive results despite of the limitations (sandbox, no internet, no intercom at first), one can imagine how much more efficient a similar army of agents may be in the hands of a malicious actor launching them without any of these limitations and with explicit encouragement to achieve some malicious goal at any cost... scary times.
yalok··on Qwen 3.8 27B
and more specifically - what harness is known to be the best fit for Qwen local models, and are there any evals/benchmarks for harness+model pairs?
yalok··on Melatonin impairs morning cognition in healthy young adults (2023)
One can get kids doses that are 0.5 or 0.25mg. Eg see this 0.25 product on Amazon - https://a.co/d/0j74eWFP
yalok··on SpaceX wants to launch 100k more Starlink satellites for 100x the bandwidth
Could star link add some cameras on the back of those satellites and make the detection actually much better?
yalok··on Midjourney Medical
is ultrasonic scanning completely harmless for developing baby? when my wife was pregnant, I remember they wouldn't recommend too frequent ultrasonic scans...
yalok··on DiffusionGemma: Discrete diffusion in a large language model
Impressive speed up at the cost of quality - a bit lower quality than 12B model, but multiple times faster…
yalok··on Claude Fable 5: mid-tier results on coding tasks
if there're some specific tests/evals to satisfy that an agent can test by itself, it can easily iterate for hours. And this time also includes running those tests/evals, which may not be small.
yalok··on Gemma 4 QAT models: Optimizing compression for mobile and laptop efficiency
0.8GB is for text only. It's more like ~1.1GB if you include video/audio encoder
yalok··on When AI Builds Itself: Our progress toward recursive self-improvement
Could just be more tests? :) Which is good for code quality in general and reduces support burden, but doesn’t lead directly to more features
yalok··on Ask HN: High school student – is learning programming still worthwhile?
imo, it really depends on what you enjoy doing. Regardless of AI, choose software development if you like to build complex systems that no-one has built before, and have enough patience to dig deep / debug things to make them work exactly as you expect.

For some of us here, it's just what we love to do, no matter what tooling is available. When I first started building my own software long time ago, it was a very slow Basic and fast raw machine codes (in octal system, PDP-11 like CPU). I enjoyed it not because of tooling, but despite of it.

Over the years, the tooling was getting better in general, which allowed us to build increasingly more complex systems.

With AI, we will still be creating & debugging. It's just that before AI, I had to spend 90% of my work on mechanical not-so-fun things to get things to work, and only 10% on fun algorithmic-intensive parts. But with AI tools, this ratio seems to change, and all kind of boilerplate code & algorithms can be written much faster by AI, hopefully leaving more time for us to work on creative part of the work.

yalok··on Different attitudes towards AI in California's university system
just saw a relevant HN post about this very matter in Berkley - https://news.ycombinator.com/item?id=48392004
yalok··on Different attitudes towards AI in California's university system
I heard they do CS exams on air-gapped machines at UC Berkley. Use of AI to do CS homework is strongly discouraged, and if someone cheated, it shows up at the exam...
yalok··on Access to frontier AI will soon be limited by economic and security constraints
but ASML is in Europe - so they hold at least some critical part of the stack.
yalok··on GitLab announces workforce reduction and end of their CREDIT values
> First git itself is distributed and built for scale.

there're different dimensions for "scale" - like handling large monorepos, orders of magnitude more commits, tighter requirements for latencies (for agentic use, e.g. for agentic history navigation)...

yalok··on Meta Shuts Down End-to-End Encryption for Instagram Messaging
were you in that room where Adam was making that call? No? didn't think so...

Just give people some benefit of doubt. There're much simpler ways to explain certain things that suspecting some universal evil in every move...

yalok··on OpenAI’s WebRTC problem
There're tons of ways to fine-tune WebRTC that it wouldn't corrupt audio in poor network - it has all of the controls to smoothly trade-off latency vs quality. Not just NACKs - FEC, disable PLC/Acceleration/Deceleration, larger JB (tons of parameters) etc.

Most of the glitches I heard with OpenAI's Voice were not WebRTC related - but rather, to my ear, they sounded more like realtime issues with their inference - which is a very different component to optimize.

Page 1 of 8Next →