trained and made a viz for the model and then made it displace text.
should probably do a proper write-up:https://x.com/i/status/2038367016969724259
289 karma · joined July 13, 2023
trained and made a viz for the model and then made it displace text.
should probably do a proper write-up:https://x.com/i/status/2038367016969724259
i feel that sometimes a lot of the layers might just be redundant and are not fully needed once a model is trained.
It works like a run club, where you have to make a review first to see other people's reviews.
I am currently implementing watchlists, comments and a mural to make it feel a bit less lonely. Right now I like the UI but it feels to lonely.
I recently made LLMs play Minesweeper and ALL LLMs that I tested had a pretty bad win to loose ratio. Like the only model that won more than 3 times was R1 (mind you there were 50 games).
[1] https://snats.xyz/pages/articles/optimizing_images.html [2] https://arxiv.org/pdf/2106.14843
[1] https://snats.xyz/pages/articles/from_bigram_to_infinigram.h...
[1] https://arxiv.org/pdf/2308.15996v1.pdf [2] https://www.adept.ai/blog/fuyu-8b [3] If you don't want to search for the figures I made a tiny post about it on my weblog: https://weblog.snats.xyz/posts/2024/02/16/
This method is better than Chain of Thought and I think is a step in the right direction.
I'll try later and post results.
I used a tiny embeddings model and PCA for dimensionality reduction.
[1] https://herman.bearblog.dev/the-chatgpt-vs-bear-blog-spam-wa...