HNHacker News
TopNewBestAskShowJobs

danielhanchen

3,617 karma · joined September 8, 2021

Unsloth github.com/unslothai/unsloth - finetune Llama 2x faster + use 70% less VRAM

1. Used to work at NVIDIA RAPIDS cuML

2. Discord: https://discord.gg/unsloth

3. Github: https://github.com/danielhanchen

4. Twitter / X: x.com/danielhanchen

5. Email: my handle @ gmail.com

6. Bug fixes for Gemma: https://news.ycombinator.com/item?id=39671146

7. Bug fixes for Gradient Accumulation: https://x.com/danielhanchen/status/1846235913443262891?lang=en

submissionscomments

Mistral Medium 3.5 YaRN bug fix

huggingface.co·1 pts·danielhanchen·
0

Gemma 4 Fine-Tuning Guide

unsloth.ai·2 pts·danielhanchen·
0

Show HN: Unsloth Studio - Local Fine-tuning, Chat UI

github.com·8 pts·danielhanchen·
2

Qwen3.5: Towards Native Multimodal Agents

qwen.ai·434 pts·danielhanchen·
214

Qwen3-Coder-Next

qwen.ai·735 pts·danielhanchen·
429

Qwen-Image-2512

qwen.ai·7 pts·danielhanchen·
1

Kimi K2 Thinking: How to Run Locally

docs.unsloth.ai·3 pts·danielhanchen·
0

LoRA Without Regret

thinkingmachines.ai·24 pts·danielhanchen·
0

Long context GPT-OSS fine-tuning

unsloth.ai·4 pts·danielhanchen·
1

Show HN: GPT OSS: How to run and fine-tune

docs.unsloth.ai·2 pts·danielhanchen·
0

Qwen3-30B-A3B-Instruct-2507

huggingface.co·5 pts·danielhanchen·
0

Qwen3-Coder: Agentic coding in the world

qwenlm.github.io·765 pts·danielhanchen·
366

2.71bit DeepSeek-V3-0324

unsloth.ai·1 pts·danielhanchen·
1

Gemma 3: Google's new multimodal models

ai.google.dev·4 pts·danielhanchen·
2

How to Run QwQ-32B effectively

docs.unsloth.ai·4 pts·danielhanchen·
3

Train your own R1 reasoning model

unsloth.ai·11 pts·danielhanchen·
5

How to run 1.58bit DeepSeek R1 with Open WebUI

docs.openwebui.com·37 pts·danielhanchen·
9

Phi-4 Bug Fixes

unsloth.ai·193 pts·danielhanchen·
68

My take on the Post Pretraining world

twitter.com·1 pts·danielhanchen·
3

Dynamic 4bit Quantization

unsloth.ai·3 pts·danielhanchen·
5

Show HN: Finetune Llama 3.2 Vision in a Colab

colab.research.google.com·10 pts·danielhanchen·
0

Python 3.11 is 1.25x faster than 3.10

docs.python.org·3 pts·danielhanchen·
5

Fixing Gradient Accumulation

huggingface.co·2 pts·danielhanchen·
0

Unit Economics of LLM APIs

lesswrong.com·5 pts·danielhanchen·
4

LoRA Learns Less and Forgets Less Updated

openreview.net·1 pts·danielhanchen·
1

VLLM automatic prefix / prompt caching

docs.vllm.ai·2 pts·danielhanchen·
1

Higher Temperatures and Min_p Sampling

arxiv.org·1 pts·danielhanchen·
1

Show HN: Open-source fine-tuning in a Colab notebook

colab.research.google.com·5 pts·danielhanchen·
0

Sahm rule signals start of recession

fred.stlouisfed.org·4 pts·danielhanchen·
3

Low Level Technicals of LLMs [video]

youtube.com·1 pts·danielhanchen·
1
Page 1 of 2Next →