HNHacker News
TopNewBestAskShowJobs

operator-name

1,084 karma · joined July 24, 2019

submissionscomments
operator-name··on Nvidia's Chat with RTX is an AI chatbot that runs locally on your PC
For some reason it's actually bundling both LLaMA 13b (24.5GB) and Ministral 7b (13.6GB), but only installed Ministral 7b. I have a 3070ti 8GB, so maybe it installs the other one if you have more VRAM?
operator-name··on Nvidia's Chat with RTX is an AI chatbot that runs locally on your PC
Having installed this, this is an incredibly then wrapper around the following github repos:

https://github.com/NVIDIA/trt-llm-rag-windows https://github.com/NVIDIA/TensorRT-LLM

It's quite a thin wrapper around putting both projects into %LocalAppData%, along with a miniconda environment with the correct dependnancies installed. Also for some reason the LLaMA 13b (24.5GB) and Ministral 7b (13.6GB) but only installed Ministral?

Ministral 7b runs about as accurate as I remeber, but responses are faster than I can read. This seems at the cost of context and variance/temperature - although it's a chat interface the implementation doesn't seem to take into account previous questions or answers. Asking it the same question also gives the same answer.

The RAG (llamaindex) is okay, but a little suspect. The installation comes with a default folder dataset, containing text files of nvidia marketing materials. When I tried asking questions about the files, it often cites the wrong file even if it gave the right answer.

operator-name··on Nvidia's Chat with RTX is an AI chatbot that runs locally on your PC
The source is available, minus the installer. You could always use the base repo after verifying it:

https://github.com/NVIDIA/trt-llm-rag-windows

operator-name··on Nvidia's Chat with RTX is an AI chatbot that runs locally on your PC
Yeah, seems a bit odd because the TensorRT-LLM repo lists Turing as supported architecture.

https://github.com/NVIDIA/TensorRT-LLM?tab=readme-ov-file#pr...

operator-name··on Nvidia's Chat with RTX is an AI chatbot that runs locally on your PC
I found some official benchmarks for enterprise GPUs, but no comparison data. I couldn't find any benchmarks for commercial GPUs.

https://nvidia.github.io/TensorRT-LLM/performance.html

operator-name··on Nvidia's Chat with RTX is an AI chatbot that runs locally on your PC
Last time I used it, LM Studio doesn't include RAG.
operator-name··on Nvidia's Chat with RTX is an AI chatbot that runs locally on your PC
This looks quite cool! It's basically a tech demo for TensorRT-LLM, a framework that amongst other things optimises inference time for LLMs on Nvidia cards. Their base repo supports quite a few models.

Previously there was TensorRT for Stable Diffusion[1], which provided pretty drastic performance improvements[2] at the cost of customisation. I don't forsee this being as big of a problem with LLMs as they are used "as is" and augmented with RAG or prompting techniques.

[1]: https://github.com/NVIDIA/Stable-Diffusion-WebUI-TensorRT [2]: https://reddit.com/r/StableDiffusion/comments/17bj6ol/hows_y...

operator-name··on Nvidia's Chat with RTX is an AI chatbot that runs locally on your PC
This is a tech demo for TensorRT, which is ment to greatly improve inference time for compatible models.
operator-name··on Spreadsheet "breaks" Apple Vision Pro eye-tracking
Honestly, this is a really cool demonstration of how their foveated rendering works.
operator-name··on Bard is now Gemini, and we’re rolling out a mobile app and Gemini Advanced
It seems to come with Google One (2TB), so with that factored in its actually quite competitive.
operator-name··on Ask HN: GitHub Requires 2FA – What For?
When you registered for 2FA it will have given you backup tokens to be written down. They're designed for this exact situation.

Given the number of software projects that use github as their canonical distribution platform and the number of supply chain attacks due to hacks, it's no pretty obvious why they've started enforcing 2FA.

operator-name··on Browsercraft – Unmodified Minecraft 1.2.5 in the browser
Note that cheerpj is not open source: https://labs.leaningtech.com/cheerpj3/licensing
operator-name··on Improving Interoperability Between Rust and C++
Announcement from the Rust side: https://news.ycombinator.com/item?id=39263750
operator-name··on The Framework 16 Teardown [video]
It's interesting that they use liquid metal for the CPU thermal interface. Unlike normal liquid metal it seems to be solid at room temperature, which I'm guessing was for ease of (re)application.

> We’re using Coollaboratory’s Liquid MetalPad through their Taiwan-based partner CCHUAN. (https://frame.work/gb/en/blog/framework-laptop-16-deep-dive-...)

It looks like the part is sold with the surrounding protective sheet: https://frame.work/gb/en/products/16-liquid-thermal-pad-amd-...

operator-name··on Google's .meme domain is here to serve your wackiest websites
Gandi lists lol.meme as a premium domain for over $800. We're also in the early access period where even normal domains cost thousands of dollars to reserve.
operator-name··on Oregon decriminalized hard drugs – it isn't working
https://jarv.is/notes/cloudflare-dns-archive-is-blocked/
operator-name··on 2023 Solar Eclipse – Oct 14
There is an equally good solar eclipse in April 2024 that crosses the country the other way around.

https://science.nasa.gov/eclipses/future-eclipses/eclipse-20...

operator-name··on This is a black and white photo, but an artist has drawn color lines through it
This is complete speculation, but the luminance difference between the red sprites and greyscale shirt is greater than that of the other colors.
operator-name··on Homebrew Website Club
This is beautiful. The simple paragraph accompanied with links to "The Dictionary of Obscure Sorrows" so eloquently expresses what I've been mulling on.
operator-name··on North Korean campaign targeting security researchers
That was the decoy behind the secondary infection vector, not the motivation.
operator-name··on North Korean campaign targeting security researchers
> The shellcode used in this exploit is constructed in a similar manner to shellcode observed in previous North Korean exploits.

At minimum the payload.

operator-name··on No way in or out of Burning Man after storm
Imgur mirror: https://i.imgur.com/TgfKL9a.jpg
operator-name··on iFixit – Why McDonald's Ice Cream Machines Are Always Broken and How to Fix Them [video]
Same topic, but a news article: https://news.ycombinator.com/item?id=37311239
operator-name··on Cheems, the Shiba Inu meme dog, has died
https://knowyourmeme.com/memes/cheems
operator-name··on Tor Release 0.4.8.1-Alpha Includes Onion Service Proof-of-Work
0.4.8.3-rc released 5 days ago, and 0.4.8 stable is set to release in the coming week or so.

https://forum.torproject.org/t/release-candidate-0-4-8-3-rc/...

operator-name··on mCaptcha: Open-source proof-of-work captcha for websites
I've not looked at the specific implementation here, but Tor's implementation[0] includes a dynamic difficulty scaling.

When under attack, legitimate users will experience a moderate delay whilst attackers will need to scale their compute.

[0]: https://gitlab.torproject.org/tpo/core/torspec/-/blob/main/p...

operator-name··on The infamous coin toss
I think I've reached an understanding, but do correct me if I'm wrong.

    lim players -> inf
    mean wealth -> starting_wealth * (win_chance * (1 + win_gain) + lose_chance * (1 - lose_loss)^rounds)
So in their example case

    win_change = lose_change = 0.5
    win_gain = 1.5
    lose_loss = 0.4
    (0.5 * (1 + 0.5) + 0.5 * (1 - 0.4)) = 1.05
so on average the mean wealth increases as the number of rounds increases. Yet at the same time for a single player

    lim rounds -> inf
    wealth -> 0
As others have mentioned, this is because the win multiplier is 1 + win_gain = 1.5 whereas the loss multiplier is 1 - lose_loss = 0.6. This is more easily seen by comparing the log of the multiplier

ln(1.5) = 0.41 ln(0.6) = -0.51

For the game to be profitable, it must satisfy

ln(1 + win_gain) > -ln(1 - lose_loss) <=> 1 + win_gain > 1/(1 - lose_loss)

So for a gain of 50% the loss must be less than 33.3%.

operator-name··on The infamous coin toss
For some reason this is not mentioned in the article but is mentioned in the video.
operator-name··on Reddit Is Down
https://www.redditstatus.com/
operator-name··on Universal Paperclips
If you're starting this on a mobile device, don't make the same mistake I did. It gets quite laggy later on and some of the "gameplay elements" are designed around having a mouse.
← PreviousPage 3 of 13Next →