HNHacker News
TopNewBestAskShowJobs

mnk47

341 karma · joined September 14, 2021

submissionscomments

Adding Error Bars to Evals: A Statistical Approach to Language Model Evaluations

arxiv.org·2 pts·mnk47·
0

Leveraging Large Language Models for Advanced Multilingual Text-to-Speech

arxiv.org·1 pts·mnk47·
1

Hertz-dev, the first open-source base model for conversational audio

si.inc·296 pts·mnk47·
56

Thinking LLMs: General Instruction Following with Thought Generation

arxiv.org·2 pts·mnk47·
1

What's the Magic Word? A Control Theory of LLM Prompting

arxiv.org·1 pts·mnk47·
0

Swarm, a new agent framework by OpenAI

github.com·258 pts·mnk47·
106

Ask HN: Why is .NET never talked about as an option for solo/small team dev?

53 pts·mnk47·
73

Recursive Introspection: Teaching Language Model Agents How to Self-Improve

arxiv.org·5 pts·mnk47·
0

The rise–and fall–of the software developer

adpri.org·3 pts·mnk47·
1

Building with OpenAI What's Ahead [video]

vimeo.com·1 pts·mnk47·
0

Chain of Thoughtlessness: An Analysis of Cot in Planning

arxiv.org·2 pts·mnk47·
0

GPT-4 users, how are you using it? Is ChatGPT Plus still worth it?

4 pts·mnk47·
5

Can Large Language Models Reason and Plan?

arxiv.org·4 pts·mnk47·
1

Wu's Method Can Boost AlphaGeometry to Outperform Gold Medalists at IMO Geometry

arxiv.org·7 pts·mnk47·
1

Ask HN: Usefulness of formal verification (Coq) and formal specification (TLA+)?

3 pts·mnk47·
1

After 2 Weeks of Testing, What Do Developers Think About Claude 3?

favtutor.com·3 pts·mnk47·
0

Ask HN: Anyone here using Gemini Pro 1.5? How does it compare to Claude Opus?

1 pts·mnk47·
1