Oh neat, I'm working on the same thing with python and playwright.
I'm finding that the latency with web LLMs is a pain in the ass and hoping to switch to llama3 after I get that set up with function calling.
I'm finding that the latency with web LLMs is a pain in the ass and hoping to switch to llama3 after I get that set up with function calling.
I've tried using it with Groq's Llama 3 70B and it worked well :)