Have you worked with any Node.js projects before? I'd actually say this is a relatively sparse list of dependencies for a user-facing tool.
20 karma · joined January 2, 2011
Have you worked with any Node.js projects before? I'd actually say this is a relatively sparse list of dependencies for a user-facing tool.
"We weren't really sure how to price it, so we're using the beta period for now to figure out what mix of models people are using and trying to figure out reasonable pricing based on that, and also ironing out various bugs and sharp edges. Then we'll start charging for it; personally I'd prefer to have it be usage-based pricing rather than the monthly subscriptions that ChatGPT and Claude use, so that you can treat it more like API access for those companies and don't have to worry about message caps."
Open to feedback here! :)
- Billy
It means that we can spin up a single model server and use it for multiple people, effectively splitting the cost. Whereas if you try to rent the GPUs yourself on something like Runpod, you'll end up paying much more since you're the only person using the model.
- Billy
Thanks for testing! :)
- Billy
Sign up should be working again! Thanks for testing! :)
- Billy
We're working on fleshing out ToS, privacy policy, and company specifics, but just to answer your first question, I'm Billy Cao, an ex-Google eng, and Matt Baker is ex-Airbnb, ex-Meta.
Re: concerns, our infra will scale relatively well (several qps per model, probably), but we're still in the stages of fleshing things out and getting feedback. :)
Feel free to drop us a line at hi@glhf.chat if you wanted to chat specifics!
- Billy
We currently use vllm under the hood and vllm doesn't support Codestral (yet). We're working on expanding our model support. Hence (almost) any model.
Thanks for testing! :)
https://github.com/vllm-project/vllm/issues/6479
- Billy :)
Great point. Right now we don't log or store any chat messages for the API (only what models people are choosing to run). We do store messages for the web UI chat history and only share it with inference providers (currently together.ai) per request for popular models, but I know some hand-waved details from an HN comment doesn't suffice.
We'll get on that ASAP. :)
We do have API support! We expose an OpenAI compatible API. You can see details when logged in at https://glhf.chat/users/settings/api
Just like our web UI it supports feeding in any huggingface user/repo.
(Also available via the user menu)
Let us know if you have any questions/feedback!
That said, awesome app. It's certainly a problem I'm sure many are constantly faced with. (I know I am) My first impressions though include a lack of detailed info (how does it work?) compounded by the inconvenience of making an account for the site. (I still haven't registered)
Edit: As for a small change you can do right away, I feel like users would consider "John Doe" <jdoe@example.com> more intuitive than "Invite Name" <Invite Email>. (Consider making the name optional altogether or just having a more intuitive input method, like multiple <input> prompts)
Some strokes have huge lag during spikes, others that I made on my screen are never displayed on the other tab.
I'm certain they use this and several other techniques such that each user reflects a far less impact on Dropbox's storage than the 50GB bought, but if that's printing money than Amazon would be a first world country by now.
Many suggest that the lack of funds forces a startup to be lean and mean, but I can see how appropriating funds has the obvious benefit of allowing you to estimate and plan your risks better.