HNHacker News
TopNewBestAskShowJobs

convexstrictly

1,147 karma · joined March 26, 2023

submissionscomments

Retire the Abstractions

hazyresearch.stanford.edu·72 pts·convexstrictly·
64

Gigatoken: Fastest Tokenizer

twitter.com·26 pts·convexstrictly·
5

Gemini Flash 2.0 Thinking Experimental

github.com·4 pts·convexstrictly·
3

What Questions Are in the Chinese College Entrance Exam?

cherylwu.substack.com·1 pts·convexstrictly·
0

Generative AI Is Not Going to Build Your Engineering Team for You

stackoverflow.blog·1 pts·convexstrictly·
0

Building GPT2o – Part 1: Audio

medium.com·3 pts·convexstrictly·
1

The Geometry of Categorical and Hierarchical Concepts in Large Language Models

arxiv.org·7 pts·convexstrictly·
0

OpenAI says it has begun training a new flagship A.I. model

nytimes.com·8 pts·convexstrictly·
0

California residents: call your legislators about AI bill SB 1047

twitter.com·22 pts·convexstrictly·
11

LISA: Layerwise Importance Sampling for Memory-Efficient LLM Fine-Tuning

arxiv.org·3 pts·convexstrictly·
1

NTIA AI Open Model Weights RFC

regulations.gov·1 pts·convexstrictly·
1

Mechanics of Next Token Prediction with Self-Attention

arxiv.org·1 pts·convexstrictly·
0

Dive Deeper into Yi-9B

huggingface.co·1 pts·convexstrictly·
0

You can now train a 70B language model at home

answer.ai·4 pts·convexstrictly·
1

Shape Suffixes – Good Coding Style (2024)

medium.com·2 pts·convexstrictly·
0

Star Trek prompt optimal for grade school math on Llama-70B

twitter.com·2 pts·convexstrictly·
1

(US Dept of Commerce) NTIA Solicits Comments on Open-Weight AI Models

commerce.gov·1 pts·convexstrictly·
0

BitDelta: Your Fine-Tune May Only Be Worth One Bit

arxiv.org·2 pts·convexstrictly·
2

Time is encoded in the weights of finetuned language models

arxiv.org·124 pts·convexstrictly·
55

Zoology 1: Measuring and Improving Recall in Efficient Language Models

hazyresearch.stanford.edu·2 pts·convexstrictly·
1

TinyGSM: Achieving >80% on GSM8k with small language models

arxiv.org·2 pts·convexstrictly·
1

Androids built to meet the labor demands

1x.tech·1 pts·convexstrictly·
1

Sam Altman likely to start company with researchers from OpenAI: Bloomberg

twitter.com·4 pts·convexstrictly·
6

Three senior researchers have resigned from OpenAI

879 pts·convexstrictly·
672

Ron Conway strongly disapproves of Sam Altman's firing

twitter.com·1 pts·convexstrictly·
0

Sutskever: OpenAI board doing its mission to build AGI that benefits all

twitter.com·121 pts·convexstrictly·
172

Kara Swisher: OpenAI dev day and store were "pushing too fast

twitter.com·2 pts·convexstrictly·
0

GPT4 coding regression claims misleading

twitter.com·4 pts·convexstrictly·
0

Model 4 bit inference 4.2x faster than 16 bit with full HF support

twitter.com·2 pts·convexstrictly·
1

SqueezeLLM: Dense-and-Sparse Quantization

arxiv.org·5 pts·convexstrictly·
1
Page 1 of 2Next →