HNHacker News
TopNewBestAskShowJobs

zagwdt

185 karma · joined September 29, 2023

submissionscomments

Score Centering Stabilizes Off-Policy Reinforcement Learning

arxiv.org·3 pts·zagwdt·
0

A/B testing LLMs in production

together.ai·2 pts·zagwdt·
0

Inference Optimization for MiniMax Sparse Attention

together.ai·1 pts·zagwdt·
0

DeepSeek V4 in vLLM: Efficient Long-Context Attention

vllm-website-pdzeaspbm-inferact-inc.vercel.app·3 pts·zagwdt·
0

Introspective Diffusion Language Models

introspective-diffusion.github.io·281 pts·zagwdt·
55

EinsteinArena: Harnessing the collective intelligence of agents in the wild

einsteinarena.com·5 pts·zagwdt·
0

RL Meets Adaptive Speculative Training

together.ai·2 pts·zagwdt·
0

Weak models excel at long context tasks

together.ai·2 pts·zagwdt·
0

TorchSpec: Speculative Decoding Training at Scale

pytorch.org·2 pts·zagwdt·
0

Flash Attention 4

together.ai·1 pts·zagwdt·
0

CoderForge-Preview: SOTA open dataset for training efficient coding agents

together.ai·1 pts·zagwdt·
0

Two years of vector search at Notion: 10x scale, 1/10th cost

notion.com·2 pts·zagwdt·
0

Consistency diffusion language models: Up to 14x faster, no quality loss

together.ai·219 pts·zagwdt·
96