15 karma · joined August 11, 2022
(found of sid.ai so obv biased)
It boggles my mind how you can train a frontier model but not write a tweet without an obvious typo.
We essentially let the model learn to retrieve like a human would: Make a first search, read the results, and then make another. This lets the model be vastly better than pre-programmed pipelines. We test this extensively and compare against implementing this with API models (like Sonnet 4.5 and GPT-5.1). SID-1 compares favorably.
Happy to answer any questions or get feedback. First and foremost: Enjoy the read. It's much more detailed than most tech reports.
I can get (even more) customization by using pandas etc., but it's usually much slower and you get much less of an intuition about the data.