220 karma · joined November 11, 2016
>In s1, when the LLM tries to stop thinking with "</think>", they force it to keep going by replacing it with "Wait".
>>I'd caution against posting a laptop to Bangladesh as there is a very real chance it'd spend the next six months sat in customs while an official looks for a pay off.
>Yeah, this is the reason I wasn't too keen on having stuff shipped from outside.
[1] https://github.com/OXY2DEV/markview.nvim/issues/218#issuecom...
[1] https://www.oreilly.com/library/view/ray-tracing-from/978149...
[1] https://www.edpb.europa.eu/news/news/2023/12-billion-euro-fi...
[2] https://ec.europa.eu/commission/presscorner/detail/en/ip_24_...
3.9% of Barcelona's GDP in 2022.
Source: https://www.statista.com/statistics/1346730/tourism-contribu...
5.4% of Catalonia's GDP in 2022.
Source: https://economia.gencat.cat/web/.content/70_economia_catalan...
11.6% of Spain's GDP in 2022.
Source: https://www.bde.es/f/webbe/GAP/Secciones/SalaPrensa/Interven...
Definitely not, Texas is ~700,000 km2 while Spain is ~500,000 km2.
The US movies and television industry can rarely get any language right. This reminds me of the Dutch-speaking scene in Oppenheimer which no Dutch speaker can understand [1]. It's like a sped-up pronunciation of a nonsense Dutch sentence by an English-speaker trying to speak a horrible German accent.
Maybe it's not feasible for you but if you're ever near an Apple Store, you could definitely call them and ask whether they have an accessibility expert you could talk to. In the Barcelona Apple Store for example, there is a blind employee who is an expert at using Apple's accessibility features on all their devices. He loves explaining his tips and tricks to anyone who needs to.
Someone did this on Reddit 9 years ago. It was remarkably good.
https://old.reddit.com/r/SubredditSimulator/comments/3g9ioz/
>Are you GPT-4.5 or GPT-5?
>I'm based on GPT-4. There isn't a "GPT-4.5" or "GPT-5" version specific to my model. If you have any questions or need assistance, feel free to ask! I'm here to help with information up to my last update in November 2023. What would you like to know or discuss?
At most, your project is a simple language model, but definitely not a large language model.
After looking at your code, you also seem to have a wrong understanding of embeddings. Other people in this thread have already shared great resources that offer good explanations on these topics.
You seem to be very dismissive of Markov chains (or HMMs) while they've been used in NLP for years and produce really great results. Your project seems to use some concepts of Markov chains but more simplified. Here [1] is a random article that does next-token prediction using Markov chains. It's basically what your project does but in a more scalable and probabilistic way.
>Use cases include: Auto-completion, auto-correct, spell checking, search/lookup, conversation simulation (chatbot), and more.
These use cases are not possible using the concepts used in your project.
[1] https://bespoyasov.me/blog/text-generation-with-markov-chain...
The paragraph in question, for posterity:
>The artwork was stolen yet again in 1934 by stockbroker and businessman Arsène Goedertie. He sent a group of men to the cathedral one evening to steal the bottom left panel known as The Just Judges. Goedertie had sliced the panel in half, leaving the other half for the police to find, and demanded a ransom for the panel. The Belgian Minister refused, so the panel is still missing to this day. You can still check out the entire 12-panel Ghent Altarpiece on display, with an impressive replica of the missing panel.