HNHacker News
TopNewBestAskShowJobs

axiom92

550 karma · joined February 8, 2015

https://madaan.github.io
submissionscomments
axiom92··on Muse Spark 1.1
What do you think about Grok 4.5 in comparison to Muse Spark 1.1?
axiom92··on Program-of-Thought Prompting Outperforms Chain-of-Thought by 15% (2022)
https://www.reddit.com/r/ChatGPT/comments/14sqcg8/anyone_els...
axiom92··on Program-of-Thought Prompting Outperforms Chain-of-Thought by 15% (2022)
This was integrated in gpt4 2 years ago:

https://www.reddit.com/r/ChatGPT/comments/14sqcg8/anyone_els...

axiom92··on Program-of-Thought Prompting Outperforms Chain-of-Thought by 15% (2022)
And even before this work, there was "PAL: Program-aided Language Models" (https://arxiv.org/abs/2211.10435, https://reasonwithpal.com/).

Afaik PaLM (Google's OG big models) tried this trick, but it didn't work for them. I think it's because PaL used descriptive inline comments + meaningful variable names. Compare the following:

```python

# calculate the remaining apples

apples_left = apples_bought - apples_eaten

```

vs.

```python

x = y - z

```

We have ablations in https://arxiv.org/abs/2211.10435 showing that both are indeed useful (see "Crafting prompts for PAL").

axiom92··on ChatGPT Atlas
You can do this at grok.com.

There is a "start thread" option below every conversation. You can also read the responses aloud (helpful if you want to do something async).

axiom92··on BERT is just a single text diffusion step
Yeah, that's the first formal reference I remember as well (although, BERT is probably the first thing NLP folks will think of after reading about diffusion).

I collected a few other text-diffusion early references here about 3 years ago: https://github.com/madaan/minimal-text-diffusion?tab=readme-....

axiom92··on Adaptive LLM routing under budget constraints
From last neurips https://automix-llm.github.io/automix/
axiom92··on Grok 4 Launch [video]
The demo was done live (as was everything else).
axiom92··on The Dark Forest hypothesis is absurd
Looks pretty cool https://www.youtube.com/watch?v=mogSbMD6EcY

Although, it seems it's only going to cover the first book (which makes sense, given how difficult the other two would be to film). The real magic for me was in book 3. It was inspiring to see someone think so far out, so boldly.

axiom92··on Turing Complete Transformers: Two Transformers Are More Powerful Than One
Welcome to one of the most hated parts of the academia.
axiom92··on Sam Altman: if I start going off, board should go after me for my shares
The joke is that he doesn't own any OpenAI shares.

[1] https://www.cnbc.com/2023/03/24/openai-ceo-sam-altman-didnt-...

axiom92··on Fuyu-8B: A multimodal architecture for AI agents
Right, but no separate image encoder + half the size could be very helpful for many applications.
axiom92··on Against LLM Maximalism
> tasks that need deterministic outputs and the thing you need to create is already known statically

Wow, interesting. Do you have any example for this?

I've realized that LLMs are fairly good at string processing tasks that a really complex regex might also do, so I can see the point in those.

axiom92··on How Is LLaMa.cpp Possible?
> And basically all servers will have 8xA100

for those wondering: no this is not the norm. My lab at CMU doesn't own any A100s (we have A6000s).

axiom92··on Never waste a midlife crisis
> The options seemed to be: If I went for it, I’d be penniless, and if I didn’t go for it, I’d be bitter. I’d be bitter going forward. Penniless certainly beats bitter. So I made the decision.

Kind of like industry -> PhD decision.

axiom92··on AI agents that “self-reflect” perform better in changing environments
Some of our recent/relevant work: https://selfrefine.info/
axiom92··on Alien Supercivilizations Absent from 100k Nearby Galaxies (2015)
https://en.wikipedia.org/wiki/Dark_forest_hypothesis

Also a major theme of https://www.amazon.com/Dark-Forest-Remembrance-Earths-Past/d...

axiom92··on Large Language Models Are Human-Level Prompt Engineers
For those curious about self-refining systems: https://selfrefine.info/ (our recent work).
axiom92··on OpenAI to discontinue support for the Codex API
cushman and code-davinci are similar for sure (same architecture). Perhaps that's what they meant.
axiom92··on OpenAI to discontinue support for the Codex API
Codex (code-davinci-002) is free (limited beta).
axiom92··on OpenAI to discontinue support for the Codex API
Actually, there is no way to be sure^. If you think about the costs + scale, it's likely to be cushman (code-cushman-001).

^ Unless you are from OpenAI, in which case I have more questions for you :)

axiom92··on OpenAI to discontinue support for the Codex API
codex (code-davinci-002) was free. This is going to be a huge deal for research groups.
axiom92··on ChatGPT has trouble giving an answer before explaining its reasoning
We did some work in exploring why spelling out the rationale before the answer works so well!

Talk: https://madaan.github.io/res/presentations/TwoToTango.pdf

Paper: https://arxiv.org/pdf/2209.07686.pdf

axiom92··on OpenAI's Foundry leaked pricing says a lot
Sure, but we don't know if ChatGPT is based on the original GPT-3 architecture.
axiom92··on Keep your AI claims in check
As they also mention, this is not the first time FTC has done this. Here is the earlier AI guidance from 4/2021: https://www.ftc.gov/business-guidance/blog/2021/04/aiming-tr....
axiom92··on The entire prompt of Microsoft Bing Chat?
Ummm not really? Any decent language model will produce sentences that look like legit instructions given this user's prompts.
axiom92··on GraphGPT: Extrapolating knowledge graphs from unstructured text
It has been possible to generate impressive graphs from text since GPT-2. Though you need a few tricks to make it work.

Here's an example (my work): https://aclanthology.org/2021.naacl-main.67.pdf

TLDR of the input/output: https://madaan.github.io/res/tldr/graph_gen_tldr.jpg

Some work from AllenAI: https://proscript.allenai.org/

axiom92··on Ask HN: Something you’ve done your whole life that you realized is wrong?
I see. I wonder if you've been phrasing this (tricky) question correctly.

For example, if you've been asking "I don't smell bad right?? I smell great, right!?" you're unlikely to get honest replies.

Reminds me of "The Mom Test" [1]. The book has a few tricks that can help with asking tough questions and receiving an honest feedback.

[1] https://www.momtestbook.com/

axiom92··on Ask HN: Something you’ve done your whole life that you realized is wrong?
> until your microbiome is able to handle all the waste / oil that your skin produces

Or maybe until you stop noticing the smell? Ask a friend perhaps.

axiom92··on Life and work of the great Visionary, Homi J. Bhabha
Wow, that's a great negative sample for any reading class--uselessly abstruse and verbose.

Took me multiple turns even with ChatGPT to simplify it:

Using desire for control may seem manageable at first, but it leads to negative consequences and people try to justify it using false reasoning and fake authority to make it seem acceptable, even though it goes against rational thinking.

Page 1 of 7Next →