Though, unlike the creators of benchmarks like Terminal Bench or ARC AGI, the Artificial Analysis Index team does not seem to have deep technical or ML backgrounds. They are ex-strategy consultants, McKinsey, et. al.
345 karma · joined June 10, 2011
harshitsurana.com
Though, unlike the creators of benchmarks like Terminal Bench or ARC AGI, the Artificial Analysis Index team does not seem to have deep technical or ML backgrounds. They are ex-strategy consultants, McKinsey, et. al.
Claude often makes better looking interfaces and designs. And I think OpenAI has solved more open math/ stats/ CS problems.
In 1934, David Hilbert, by then a grand old man of German mathematics, was dining with Bernhard Rust, the Nazi minister of education. Rust asked, “How is mathematics at Göttingen, now that it is free from the Jewish influence?” Hilbert replied, “There is no mathematics in Göttingen anymore.”
Meanwhile, this is the blog of our earliest work from January that will be published in ICML24: https://blog.allenai.org/data-driven-discovery-with-large-ge...
Also if you are interested - happy to correspond more by email. Lot more to share privately.
This is already being used by a few hundred scientists in a lab. We are aiming to extend to thousands next year with a focus on CS, climate & bio.
I am using these tools to send custom personal messages to close friends and family.
Honestly, Sam is, along with Steve Jobs, the founder I refer to most when I'm advising startups. On questions of design, I ask "What would Steve do?" but on questions of strategy or ambition I ask "What would Sama do?"
What I learned from meeting Sama is that the doctrine of the elect applies to startups. It applies way less than most people think: startup investing does not consist of trying to pick winners the way you might in a horse race. But there are a few people with such force of will that they're going to get whatever they want.
Counted 52 OpenAI employees supporting @sama with more rolling in by the minute
https://twitter.com/FreddieRaynolds/status/17261100249827369...
I was in a Ph.D. program at a top CS school and there are ways to transition your visa when building a startup. It was that I was not sure if the transition or the startup would work out - that startup did not - but years later another one did.
I would probably not have taken the plunge out of academia and not achieved much else had it not been for him. And I am deeply grateful for that.
It forever tuned me in to the ethos of Silicon Valley. And I have tried paying back where I can.
And this is besides the other stuff that search ones can’t do directly but it does quite well. It has definitely cut web searches for me.
[1] https://www.reuters.com/technology/openai-track-generate-mor...
Exactly, there is a paradox that is getting more extreme by the day - that social media (and the broader web) is a wellspring of knowledge, yet also a vortex of addiction, filter bubbles & lost productivity. I want one without the other.
This pushed me to start building open source at OpenLocus, contributions & feedback are welcome - details in other comment.
In the more medium term we are collaborating on improving information overload, filter bubbles & misinformation with labs at AllenAI, CMU, UPenn & Utah.
Contributions and feedback are welcome! Feel free to hit me up for an early access - email in the profile.
Having been involved in both traditional machine learning and common sense AI in my grad school years, I've seen first-hand the limitations of a purely statistical approach. (Some of my past data augmentation work is being used to benchmark LLM reasoning.)
While most folks are too fixated by the 'quick wins' achieved by LLMs the trade-off is often a lack of non shallow reasoning. And I worry that many active researchers are glossing over these deeply rooted issues.
Look forward to the day when such flaws are largely eliminated!