HN
Hacker News
Top
New
Best
Ask
Show
Jobs
Comment by rurban | Hacker News Reader
Full thread
rurban
·
For LLM inference in production we use vLLM (with our own model).
View on HN
idbnstra
·
did you fine tune another model, and if so which one?
rurban
·
Not really fine tuned. But the base was Qwen. We are analyzing images
Reply on news.ycombinator.com