HNHacker News
TopNewBestAskShowJobs

ed

5,971 karma · joined February 23, 2007

emcmanus -at- gmail
submissionscomments
ed··on Show HN: Llama 3.3 70B Sparse Autoencoders with API access
Early traffic laws were actually created in response to child pedestrian deaths (7000 in 1925).

https://www.bloomberg.com/news/features/2022-06-10/how-citie...

ed··on Show HN: Llama 3.3 70B Sparse Autoencoders with API access
This is the ultimate propaganda machine, no?

We’re social creatures, chatbots already act as friends and advisors for many people.

Seems like a pretty good vector for a social attack.

ed··on Stopping by Woods on a Snowy Evening (1923)
Apparently Eric Whitacre (a choral composer popular with high school choirs around 2001-2003) originally wrote Sleep to the lyrics of “Stopping by…” but was sued by Frost’s estate. He can’t release the original until 2038. https://ericwhitacre.com/music-catalog/sleep
ed··on After 3 Years, I Failed. Here's All My Startup's Code
There seems to be a subtext here of “see? I told you it was hard!”

Nothing wrong with setting out to build a Google instead of a Basecamp. They’re different kinds of company.

If anything it’s easy to underestimate the risk of building a low-risk business. They’re all hard.

Kudos to Konfig!

ed··on Ilya Sutskever NeurIPS talk [video]
Based on the context Ilya is not referring to that kind of agent. He’s referring to something more fundamental (which I was curious about, too).
ed··on Model Context Protocol
I’ve gone looking for services like this but couldn’t find much, any chance you can link to a few platforms?
ed··on Show HN: AI OmniGen – AI Image Generator with Consistent Visuals
Elegant architecture, trained from scratch, excels at image editing. This looks very interesting!

From https://arxiv.org/html/2409.11340v1

> Unlike popular diffusion models, OmniGen features a very concise structure, comprising only two main components: a VAE and a transformer model, without any additional encoders.

> OmniGen supports arbitrarily interleaved text and image inputs as conditions to guide image generation, rather than text-only or image-only conditions.

> Additionally, we incorporate several classic computer vision tasks such as human pose estimation, edge detection, and image deblurring, thereby extending the model’s capability boundaries and enhancing its proficiency in complex image generation tasks.

This enables prompts for edits like: "|image_1| Put a smile face on the note." or "The canny edge of the generated picture should look like: |image_1|"

> To train a robust unified model, we construct the first large-scale unified image generation dataset X2I, which unifies various tasks into one format.

ed··on Quantized Llama models with increased speed and a reduced memory footprint
Grammar samplers are clever! But in the case of a missing escape character you’ll end up with a corrupted string.

Take for example: "A dog says \"Woof!\""

With a grammar, you’ll end up with "A dog says " when the model forgets to escape.

Which is valid JSON, but not what the model intended.

So it’s usually better to catch the exception and ask the model to try again.

Unless you’ve come across a sampler with backtracking? That would be cool

ed··on Quantized Llama models with increased speed and a reduced memory footprint
Oh cool! I’ve been playing with quantized llama 3B for the last week. (4-bit spinquant). The code for spinquant has been public for a bit.

It’s pretty adept at most natural language tasks (“summarize this”) and performance on iPhone is usable. It’s even decent at tool once you get the chat template right.

But it struggles with json and html syntax (correctly escaping characters), and isn’t great at planning, which makes it a bad fit for most agenetic uses.

My plan was to let llama communicate with more advanced AI’s, using natural language to offload tool use to them, but very quickly llama goes rogue and starts doing things you didn’t ask it to, like trying to delete data.

Still - the progress Meta has made here is incredible and it seems we’ll have capable on-device agents in the next generation or two.

ed··on Ask HN: Website with 6^16 subpages and 80k+ daily bots
As others have pointed out the calculation is 16^6, not 6^16.

By way of example, 00-99 is 10^2 = 100

So, no, not the largest site on the web :)

ed··on Stripe to acquire Bridge for $1.1B
The vast majority of Stripe's fees go to payment processors. I am sure Stripe would prefer to collect those fees for themselves, and pass along some savings to merchants. Crypto is the best avenue for Stripe to do that.
ed··on Thinking LLMs: General Instruction Following with Thought Generation
This paper comes from Meta and introduces Thought Preference Optimization (TPO), a post-training process that encourages small models to think, similar to o1.

The results are impressive - Llama 3 8b performs almost on par with GPT-4o across a wide range of tasks, not just logic and math.

Interestingly, the post-training process significantly improves model performance even without “thoughts” (the “direct baseline” case in the paper).

ed··on Apple's New iPad Mini Highlights the Company's AI Advantage
iPad Mini’s (and iPhone’s) 8gb memory will be very limiting. It’s sufficient for 3b models quantized to 4bits, but puts significantly more powerful 8b models just out of reach.

This isn’t a huge issue now, but with another 12 months of local AI progress it’ll leave a competitive opening for device-makers who ship more memory. (See Meta’s TPO paper for significant improvements to Llama, pending release.)

I skipped this iPhone cycle despite being on Apple’s iPhone upgrade program — a memory bump to the Pro line would’ve easily justified an upgrade.

ed··on Un Ministral, Des Ministraux
> GP misunderstood

I don’t think it’s fair to claim the weights are available if you need to hammer out a custom agreement with mistral’s sales team first.

If they had a self-serve process, or some sort of shink-wrapped deal up to say 500k users, that would be great. But bespoke contracts are rarely cheap or easy to get. This comes from my experience building a bunch of custom infra for Flux1-dev, only to find I wasn’t big enough for a custom agreement, because, duh, the service doesn’t exist yet. Mistral is not BFL, but sales teams don’t like speculating on usage numbers for a product that hasn’t been released yet. Which is a bummer considering most innovation happens at a small scale initially.

ed··on Amazon reveals first color Kindle, new Kindle Scribe, and more
You might be interested in the Kindle Basic. It's the smallest in the lineup and a comparable size to the first-gen oasis (before they increased the screen size) – my previous daily carry.

It's almost a pocketbook form-factor. I overlooked it initially because who wants a basic model? but the only thing I miss in practice is waterproofing. That, and the Oasis OEM cover which was unexpectedly nice, like a leather-bound pocketbook.

ed··on Un Ministral, Des Ministraux
3b is is API-only so you won’t be able to run it on-device, which is the killer app for these smaller edge models.

I’m not opposed to licensing but “email us for a license” is a bad sign for indie developers, in my experience.

8b weights are here https://huggingface.co/mistralai/Ministral-8B-Instruct-2410

Commercial entities aren’t permitted to use or distribute 8b weights - from the agreement (which states research purposes only):

"Research Purposes": means any use of a Mistral Model, Derivative, or Output that is solely for (a) personal, scientific or academic research, and (b) for non-profit and non-commercial purposes, and not directly or indirectly connected to any commercial activities or business operations. For illustration purposes, Research Purposes does not include (1) any usage of the Mistral Model, Derivative or Output by individuals or contractors employed in or engaged by companies in the context of (a) their daily tasks, or (b) any activity (including but not limited to any testing or proof-of-concept) that is intended to generate revenue, nor (2) any Distribution by a commercial entity of the Mistral Model, Derivative or Output whether in return for payment or free of charge, in any medium or form, including but not limited to through a hosted or managed service (e.g. SaaS, cloud instances, etc.), or behind a software layer.

ed··on Longwriter – Increase llama3.1 output to 10k words
Paper: https://arxiv.org/abs/2408.07055

The model is stock llama, fine tuned with a set of long documents to encourage longer outputs.

Most of the action seems to happen in an agent.

ed··on Ask HN: Is anyone working at least 4 hours daily on an Apple Vision Pro?
It’s not great as an external display, since the experience is like sitting very close to a massive TV while wearing steampunk googles. You have to turn your head a lot, and eye-tracking is broken for a lot of sites in Safari (eg accidental downvotes on HN).

I don’t think that’s a dealbreaker - as long as the device can actually act as a standalone computer.

But it can’t, at least not for developers. It’s much closer to an iPad that only runs 10% of the apps you want. Maybe this is sufficient if your job only requires communication, and not, like, actual work.

So you end up with an expensive, socially awkward accessory to your MBP, which quickly gets left at home, because it doesn’t really do anything better than your existing devices.

(The one use case AVP handles well: watching a movie, in bed, alone. Which is kinda bleak.)

ed··on Ask HN: Good Sites for/with AI Enthusiasts?
yep!
ed··on Ask HN: Good Sites for/with AI Enthusiasts?
https://www.youtube.com/@SECourses
ed··on Ask HN: Good Sites for/with AI Enthusiasts?
I'm on:

Image

- Terminus Research Group, from bghira of SimpleTuner/diffusers https://discord.gg/cSmvcU9Me9

- AI Toolkit https://github.com/ostris/ai-toolkit https://discord.gg/VXmU2f5WEU

- Stable Diffusion https://discord.gg/stablediffusion

LLM

- LLamaIndex https://www.llamaindex.ai https://discord.com/invite/eN6D2HQ4aX

- Nous research https://discord.gg/nousresearch

- LangChain https://discord.gg/hMrfPpUk

Platforms

- Replicate https://discord.gg/replicate

- Fal https://discord.gg/fal-ai

ed··on Ask HN: Good Sites for/with AI Enthusiasts?
It’s a big field!

But if you’re in a few discords and a bunch of subreddits, you’re doing it right.

The most interesting stuff happens in GitHub PR’s, but you have to know where to look. Kohya’s misnamed SD3 branch has a ton of good flux hints, for example. It’s also where furkan gets pretty much all his content, before it gets paywalled.

Unfortunately, unless you participate full-time it’s hard to follow along. But if you really dig in and learn to modify your tooling (Comfy, kohya etc), you’ll start to come across some really impressive people who are all self-taught, and very accessible.

It’s totally possible to work your way up to the frontier with a few months of hacking. (And disposable income for GPU time.)

And the overlap between image AI’s and LLM’s is actually pretty great since they’re all transformers under the hood.

Civit, in my experience, is a good source for weights but most of the guides are written by people without much actual experience.

If you haven’t already, use tensorflow or wandb to get an intuitive understanding of your training parameters. It’s very easy to connect your tools to these services. This is by far the most helpful thing I’ve done, and something I really regret not doing sooner.

ed··on Show HN: I made crowdwave – imagine Twitter/Reddit but every post is a voicemail
You should add a talkboy[1] inspired voice changer to make it less intimidating to leave a message

[1] - https://en.wikipedia.org/wiki/Talkboy

ed··on g1: Using Llama-3.1 70B on Groq to create o1-like reasoning chains
FYI this is just a system prompt and not a fine-tuned model
ed··on Terence Tao on O1
This. I’ve been using elixir for ~6 months (guided by Claude) and probably couldn’t solve fizz buzz at a whiteboard without making a syntax error. Eek.
ed··on Reclaim the Stack
(Except for Postgres, since Fly's solution isn't managed)

Heroku's price is a persistent annoyance for every startup that uses it.

Rebuilding Heroku's stack is an attractive problem (evidenced by the graveyard of Heroku clones on Github). There's a clear KPI ($), Salesforce's pricing feels wrong out of principle, and engineering is all about efficiency!

Unfortunately, it's also an iceberg problem. And while infrastructure is not "hard" in the comp-sci sense, custom infra always creates work when your time would be better spent elsewhere.

ed··on The Lurker's Guide to Babylon5
Interesting link! But rebooting TOS seems even lazier than Star Trek: A New Ship.

The universe is a big place; a decent writer can find a way to tell any story they want. (As demonstrated by the many IP-friendly reboots since 2004, when this was written.)

ed··on Superhuman built an engine to find product market fit (2018)
Having followed this exact playbook to validate several products I can confidently say this will give you false positives and the only reliable way to determine when you have PMF is: accidentally get PMF on something so that you know what it feels like (it’s unmistakeable).

Peter Reinhardt from Segment has a must-watch talk for anyone interested in this topic https://youtu.be/_6pl5GG8RQ4?si=pogHC45L58U7K6mW

ed··on 13ft – A site similar to 12ft.io but self-hosted
I fully agree with the sentiment! I support and do pay for sources I read frequently.

Sadly payment models are incompatible with how most people consume content – which is to read a small number of articles from a large number of sources.

ed··on Google pulls the plug on uBlock Origin
uBO's competitors do better App Store Optimization.

I used AdBlock Plus for a while because it looked like the more popular ad-blocker.

I only uninstalled AdBlock Plus because it keeps displaying a "upgrade to premium" popup.

uBO has been such an improvement that now I worry I lost some geek gred for ever using ABP.

← PreviousPage 4 of 23Next →