I realized that AIs will need to keep us around much like we keep around an old DOS box running a specific program. There are so many analogue outputs in the real world that they still will not be able to take in as well as humans can.
I ate Kiwis for years before someone told me most people DO NOT eat the skins. I looked at my bag of kiwis, and the even the damn monkey mascot was holding a spoon...mocking me...
In the US, Business' are treated like people with free speech rights. If it would be cheaper for them in the long run to use ai and robots instead of humans, they will figure out a way to make it so.
Assuming you had a 1:1 scale, what would be the difference between you taking a camera of an environment and measuring everything the same way vs. it being generated artificially in a 3d space?
Wouldn't the inputs be more or less the same wrt training a model with the synthetic data vs live video and the data captured there?
The author mentions a few apps to control window positions. I use one called Rectangle, is there any benefit the other options might have that I am missing?
Most of these are your standard botnet rings. Either accounts directly are taken over, and the attacker adds 2FA or carding rings take stolen #s and attempt to add credits.
It is...incredible how many there are. Stripe does far too little in my opinion to help prevent issues like this, even though they have the business intelligence and enough data to do so.
I built a harness from scratch prior to trying any of the ones out there, so I knew how it work in a real way. I QUICKLY understood that the biggest issue with getting my shit done is that _my_ inputs are the untrusty ones. How many time do you hit backspace in a day?
Every plan and every code checkpoint finds me saying "Check with Grok and Fable latest to critique our strategy/code review" with pretty much every model. I havent ran into any deal breakers with the new Flash version yet (like it not running a tool properly or coming back with something completely daft)
Not them, but I payed 10 dollars to DeepSeek directly to use their Reasonix tool. I worked all weekend and the past few days, billions of tokens, I still have 3 bucks left!
Hey this looks good! Maybe consider adding a hover-over popup for the rectangles explaining what each thing means to a lay person. I see it at the bottom, but that is below the fold.
Late, but it depends. Sometimes I ask for all of them to do the same thing and have a different model judge it, sometimes I ask for an orchestration model to run the smaller ones each on one task. I have a 'model cohort' where I run the same request across 5 very inexpensive models. Then have my driving LLM judge or synthesize.
You just come up with what you want in your head, and tell it to do it in that way and it does it. I trust 5 independent smartest programs ever over the SOTA smartest program's only.
I work for a router company too, I ran some tests on all of the cheapest models and came to the same outcome where a handful of small models ran together in conjunction outperform SoTA models -- outperforms in that it got a 95% vs a 94% and I bet that changes with the day of the week. Anyways, I did get a similar result in a different sort of measurement.
I've been doing a lot of Trust and Safety and anti-scam stuff recently. When it was first re-released Fable wouldn't do anything for me, now I don't really run into refusals or what it used to do, just stop returning any data at all whatsoever.
Outside topic, but check out the Inkling model -- it is SUPER FAST and does a good job at being a terminal buddy but I would probably offload the real programming or hard tasks to fable or someone else