HNHacker News
TopNewBestAskShowJobs

henry_pulver

270 karma · joined July 15, 2021

submissionscomments
henry_pulver··on Switch to Claude without starting over
Amusing that Anthropic's approach to migrating context is asking their competitor's product to hand over the data it's stored about you.

Must be some of the lowest switching costs I've seen which doesn't bode well for OpenAI's consumer revenues...

henry_pulver··on Shipmap.org
This is mesmerising!

Really well made and enjoyed the audio explanation.

It's a shame that it includes the now-mandatory discussion of how this shipping is actually bad because of carbon emissions. Seems to me the widespread availability of cheaper goods has been a great thing for humanity on balance!

henry_pulver··on Chat Control Must Be Stopped – Now
> But all coming on here and saying "ooohh, this is bad, innit!" is not very interesting, and unlikely to prevent it.

I disagree - this is how the internet can strengthen democracy.

Upvoting and commenting makes this post hit the top of HN and stay there. This makes it visible to many EU citizens who can reach out to their MEP's to ask them to vote against it. Seems a pretty effective strategy to me as someone living in a non-EU country.

Although agree that we should also be discussing the questions you raised.

henry_pulver··on Start presentations on the second slide
My understanding of the key point of this blog is:

> Instead of explaining the technical background first so listeners understand the solution to a problem, start with the problem. Then explain the context/technical background second

henry_pulver··on Map of forest sounds from around the world
Wow, this has many more recordings & a larger variety! Thanks for sharing
henry_pulver··on Ask HN: When will we hit a limit on LLM performance?
As far as I (ex-ML researcher) know, the main technological case that LLM performance will hit a limit is due to the amount of text data available to train on is limited. The ways these scaling laws work is they require 10x or 100x quantity of data to see major improvements.

This isn't necessarily going to limit it though. It's possible there are clever approaches to leverage much more data. This could either be through AI-generated data, other modalities (e.g. video) or another approach altogether.

This is quite a good accessible post on both sides of this discussion: https://www.dwarkeshpatel.com/p/will-scaling-work

henry_pulver··on Show HN: Probabilistic Tic-Tac-Toe
This is fantastic!

The dice roll animation is :chefkiss:

henry_pulver··on Ex-OpenAI employee reported losing 85% of his family's net worth
100% agreed. Must have been a very tough decision. But good for him.

Taking selfless actions like these, that have major personal costs, require serious courage.

henry_pulver··on Write OpenAPI with TypeSpec
Great idea - spend far too long reading & writing OpenAPI!

Particularly anyOf, allOf and oneOf (especially when nested) lead to really confusing nested specifications in OpenAPI. Really like how TypeSpec handles unions & intersections.

Playground is great for getting a feel for it fast too

henry_pulver··on Tokenizer UI for Mistral and Claude
On the tokenizers Mistral use for proprietary models, this isn't common knowledge.

This tokenizer is correct for the 7B open model and 8x7B MoE model. It'll probably be the closest to the ones their proprietary API-only models use

henry_pulver··on Tokenizer UI for Mistral and Claude
I use the OpenAI tokenizer UI a lot when prompt engineering.

Token count for inputs allows comparison of different data formats (YAML, JSON, TS) and is a crude measure of prompt importance weighting. For outputs it is a relative measure of output speed between prompts (tok/s varies by time of day) and a crude measure of compute used in outputs (why “Think step-by-step” works). Token count also determines the cost of a prompt.

Since there’s no equivalent for other providers, I built one for Mistral & Anthropic. If it’s useful, I can add other providers too - let me know which you’d like.

henry_pulver··on Launch HN: Glide (YC W24) – AI-assisted technical design docs
Didn't fully get the value from reading the post so thought I'd give it a try. Our company is open source so put in our actual url :)

Sadly it errored with the classic NextJS:

Application error: a client-side exception has occurred (see the browser console for more information).

The error in the console was: 601-ce9691b65ce5066e.js:4 Error: An error occurred in the Server Components render. The specific message is omitted in production builds to avoid leaking sensitive details. A digest property is included on this error instance which may provide additional details about the nature of the error.

henry_pulver··on Presenting TOC (Theory of Constraints) [video]
I read 'The Goal', which is Goldratt's allegorical book that presents this theory.

It's a little hamfisted in it's allegorizing and felt it sometimes spent a long time making simple points. Even still, it's much more memorable than this speech, so would recommend it if you have the time and inclination!

henry_pulver··on Claude for Google Sheets
Seems to me a bit of an odd launch from Anthropic

Aren't their core users developers? Why release a sheets extension?

henry_pulver··on How to Twitter Successfully
> Your guide suggested I pay for Twitter before even trying to use it

Hey, this isn't _my_ guide. Shared it because I found it interesting. :)

henry_pulver··on CarbonGPT: Ask questions about carbon emissions
I made this chatbot which has access to up-to-date emissions data through the Carbon Interface API.

Access to accurate emissions data is useful but comparative research often involves sifting through disparate sources. Combining the chatbot with the Carbon Interface API gives a quick and easy overview of CO2 emissions from a number of activities.

There are still a number of limitations. Like any LLM, it can make factual errors and mistakes when calling the API.

henry_pulver··on Great Pyramid of Cholula
Pretty sure it's this pyramid where they didn't know the whole hill the pyramid was on top of was actually part of the pyramid until recently

I wonder what other natural formations are man-made, but we don't realise it yet

henry_pulver··on Ray Kurzweil and Mitch Kapor’s “Long Bet” on the Turing Test
> How is creativity different than adding some randomness to the answer generating?

Deterministic systems are clearly not creative. But randomness is a necessary, but not sufficient, requirement for creativity.

Adding random noise to an image makes it grainy (like TV static from back in the day), it doesn't lead to new masterpieces.

henry_pulver··on Show HN: Oyie – A micro social app to stay in sync with your friends
All the positive comments seem fake
henry_pulver··on Show HN: Superflows – open-source AI Copilot for SaaS products
Not right now, but we're working on it! :)

Base Llama 2's reliability on our prompt isn't good enough right now so working on fine-tuning it to make sure it outputs the right format - we'll release this at the same time as we make it easier to self-host

henry_pulver··on Show HN: Superflows – open-source AI Copilot for SaaS products
Totally get the reliability issues you've seen - getting LLMs to output the right response 80% of the time on these kinds of tasks is not that hard, the last 20% is the real challenge.

On trying us out - ping me an email at henry@superflows.ai

henry_pulver··on Edge detection doesn’t explain line drawing
This is fascinating!

I remember reading that most optical illusions don't work on people raised in remote tribes in the Amazon, as their visual perception has been 'fine-tuned' for jungle contours, instead of the straight lines in the west.

Is it possible that we _learn_ how to perceive line drawings in our early years?

henry_pulver··on Patterns for building LLM-based systems and products
Not seen a great explainer on this yet.

You'd either need access to the model weights or a fine-tuning API.

Then depending on which fine-tuning approach you want to use, the user data you need to collect will be different: RLHF requires multiple outputs to a single query vs instruction fine-tuning where you need great input-output pairs to train on. You could ask the user's feedback after running the LLM to pick out good training data.

henry_pulver··on Pgvector 0.4.0 Performance
Had a question about the "pre-warming" technique you used - why does this help? You say it helps with RAM utilization, but why?

Does it just spin up the right number of probes up so the experiments don't include this extra cost of spinning up these probes?

(apologies if this is a dumb question - I don't know enough about vector databases to know if this is a dumb question or not!)

henry_pulver··on Britain is sleepwalking into censorship?
As a Brit, I'm embarrassed. These politicians are LARPing and haven't learned the most important political lesson of the last 100 years - that restrictions on speech are dangerous.
henry_pulver··on Why the Learning Styles myth spreads
If a theory is unfalsifiable, it's unscientific. The whole point of theories grounded in science is they allow you to make accurate predictions about the world.

What can you make predictions about if the matching hypothesis isn't true? What does this actually tell you about the world? Nothing.

henry_pulver··on Learn ML through live team competitions, not lectures
The course is "Introduction to Reinforcement Learning". We make no claims that you'll be writing papers come the end of 4 weeks! :)

Really interesting reading that Norvig's piece. I agree with almost all of it and think what we're doing at Delta Academy lines up with most of it!

Particularly:

> The key is deliberative practice: not just doing it again and again, but challenging yourself with a task that is just beyond your current ability, trying it, analyzing your performance while and after doing it, and correcting any mistakes. Then repeat. And repeat again.

This is exactly what we do - providing weekly challenges that stretch your ability (that also happen to be fun), discussing the approaches taken by teams & giving expert feedback on code.

And his recipe for programming success:

> Get interested in programming, and do some because it is fun. Make sure that it keeps being enough fun so that you will be willing to put in your ten years/10,000 hours.

> Program. The best kind of learning is learning by doing.

> Talk with other programmers; read other programs. This is more important than any book or training course.

> Work on projects with other programmers.

Again - team projects which are fun sound right up his street.

henry_pulver··on Learn ML through live team competitions, not lectures
We've designed it to be very compatible with working full time!

We suggest ~10 hours per week, although except for the 30 min live competition each week, all of these hours can be done at times that suit you.

henry_pulver··on Learn ML through live team competitions, not lectures
If it was just 1 competition, I'd agree with you.

It's more a cohort-based online class that's punctuated by competitions (once per week). The competitions serve to motivate you to learn in a fun finale each week, rather than all being about winning. :)

The price tag includes 12 tutorials with exercises, 4 competitions (incl live discussion of solutions with the cohort) and expert code review from instructors on all the exercises.

henry_pulver··on Learn ML through live team competitions, not lectures
Kaggle has big competitions over several months. They're designed to find top machine learning talent & innovative solutions to the problems they set. Typically the winners aren't just learning ML, they're seasoned pros doing it as a side project.

Our competitions are designed to teach. Each is a progression in difficulty over the previous one. Also there are a set of tutorials preceding each competition which get you up to speed on what you'll learn in the competition.

Plus we're organised into a cohort so you're not competing with the whole world - rather you're competing with your peers who are also learning. You work in a pair on the competition. Then we discuss the solutions teams came up with & what an 'ideal' solution would look like (if one exists - sometimes it doesn't!).

Page 1 of 2Next →