HNHacker News
TopNewBestAskShowJobs

advaith08

200 karma · joined August 3, 2023

meet.hn/city/us-San-Francisco

Socials:

- linkedin.com/in/advaith-sridhar

---

submissionscomments
advaith08··on Launch HN: Discovered Materials (YC P26) – AI agents to discover new materials
we have access to the thinking trace summaries, not the raw traces
advaith08··on Launch HN: Discovered Materials (YC P26) – AI agents to discover new materials
yeah, the tricky thing about the experimental loop is that: 1. its very difficult to do it in a reproducible manner (the same experiment done twice often gives different results due to small undocumented changes) 2. its expensive to do at scale.

Both of these properties make it hard to hill climb on experiment. What's worked for us so far is precisely what you said - having human experts review and provide feedback. we distil their reviews into rubrics, and have LLMs act as proxy experts using these rubrics. We expect the models will hill climb using this approach, and will reach (close to) human expert level by doing this.

advaith08··on Launch HN: Discovered Materials (YC P26) – AI agents to discover new materials
thanks for sharing! We think the Experiments-as-code path is not the right approach. the beauty of LLMs is their ability to ingest and reason over unstructured data - they remove the need for formalizing experiments. We tried using declarative templates to document our experiments, but realized that most of the interesting insights (for example, how viscous a liquid feels) is easier described by ranting about the experiment to a LLM, than formalizing it via constructs/code.
advaith08··on Launch HN: Discovered Materials (YC P26) – AI agents to discover new materials
interesting! we haven't really considered these yet. However, we may soon have to - the recurring feedback we hear from industry is that crystallinity is a pipe dream, and that most materials are going to be amorphous (or perhaps quasicrystalline). MLIPs have lowered the computation cost for a large number of atoms/odd cell size, but it may be a while before they're accurate enough to simulate these scenarios
advaith08··on Launch HN: Discovered Materials (YC P26) – AI agents to discover new materials
that's true, we've seen examples of many promising startups that are a few years in and stuck because the industry is so risk averse. will explore the other domains you've mentioned!
advaith08··on Launch HN: Discovered Materials (YC P26) – AI agents to discover new materials
good points. one of the reasons we picked the semiconductor industry is that its less price sensitive than others-companies are willing to pay if the performance is there. Effort is a different story though, and definitely a tradeoff to keep in mind. We're doing experiments ourselves now at university partner labs (UC Berkeley and Stanford), which helps us get moving quickly. At some point, we'll need a partner though - the equipment and testing process quickly get very expensive.
advaith08··on Launch HN: Discovered Materials (YC P26) – AI agents to discover new materials
yeah we were surprised by how much it does it. Our approach has been retroactive - we monitor the thinking trace, spot reward hacking behavior and then fix things. We haven't faced this issue with Sol though - its been much more well behaved
advaith08··on Launch HN: Discovered Materials (YC P26) – AI agents to discover new materials
We're still figuring this out. We'll need some synthesis equipment (think CVD, PVD etc) and characterization (XRD, Raman spectroscopy) tools in-house to validate that we're making the right materials. We're considering developing these tools in-house - the models sometimes come up with clever modifications to them so that they can deposit new materials. We think equipment is as central to new material discovery as the material itself, and will probably need to be rethought to allow for high-speed AI based experimentation
advaith08··on Launch HN: Discovered Materials (YC P26) – AI agents to discover new materials
Cool read, and agree that closing the computation > experimental loop is key!
advaith08··on Launch HN: Discovered Materials (YC P26) – AI agents to discover new materials
There’s a variety of computational techniques that help us establish some confidence on the materials. Atomistic simulations can estimate stability and bulk properties of a new material, and we have synthesis experts (min qualification: PhD in thin film deposition) come up with rubrics on how to judge if a material/synthesis recipe is worth trying. All these approaches have known limitations, and improving them is the bulk of our work as a company! There’s also a lot of work to be done in figuring out the minimal set of experiments required to know if a research direction/material set is worth pursuing
advaith08··on Notes on the new Claude analysis JavaScript code execution tool
The custom instructions to the model say:

"Please note that this is similar but not identical to the antArtifact syntax which is used for Artifacts; sorry for the ambiguity."

They seem to be apologizing to the model in the system prompt?? This is so intriguing

advaith08··on Timeline of the xz open source attack
Agree. I think a more core issue here is that only 1 person needed to be convinced in order to push malware into xz
advaith08··on Scientists have traced human tail loss to a short sequence of genetic code
The dehydrate comment is a reference to the three body problem - a great book!
advaith08··on You Shouldn't Make Friends at Work
fully agree with this. Plus, if you're worried about "promotion competitiveness" souring friendships with people in your team, you can always make friends with people in other departments and meet them during lunch. This doesnt have to be a hard binary rule
advaith08··on The Seamless Communication models
seen a lot of these, but none for Indian languages. Would love to try an Indian language one!
advaith08··on A coder considers the waning days of the craft
I think the point of the comparison with chess is to show what happens when AI becomes vastly better than humans at something. GPT4 is not vastly better than humans at coding, but the author is using chess as an analogy to visualize what the world might look like if these GPTs continue to get better
advaith08··on Ask HN: Who wants to be hired? (November 2023)
Location: Bay Area, SF Remote: No Willing to relocate: Yes Technologies: Pytorch, Python, NLP, CV, LLMs Résumé/CV: https://drive.google.com/file/d/1Ac_zraUpg4MAojQekE_XSNEfhs4... Email: advaith@cmu.edu
advaith08··on OpenAI just replaced its core values with completely different ones
agree with this. in their new values, they literally use the same words that were in the old ones ("unpretentious", "collaboration"). I also dont think this is a change "at the drop of a hat" as the article suggests - OpenAI's been through a major inflection point with GPT3.5/4, and the company's not the same one it was a few years ago. It makes sense to make updates to the company's core values
advaith08··on OpenAI is too cheap to beat
imo they dont have batching because they pack sequences before passing through the model. so a single sequence in a batch on OpenAI might have requests from multiple customers in it
advaith08··on Show HN: WhatsApp-Llama: A clone of yourself from your WhatsApp conversations
Oh my apologies, this is Llama 2! Im using Llama 2 7b chat here, not the original Llama
advaith08··on Show HN: WhatsApp-Llama: A clone of yourself from your WhatsApp conversations
yeah I think there are a lot of use cases for an assistant trained on your chat history. Given how privacy sensitive this use case is, I think maybe Apple is the best suited to build something like this? Hope they come out with something cool
advaith08··on Show HN: WhatsApp-Llama: A clone of yourself from your WhatsApp conversations
Very cool. I like the introspection bit, I've realised quite a bit about my texting style from talking to Llama too. I think Im also very "type first and think later" on WhatsApp
advaith08··on Show HN: WhatsApp-Llama: A clone of yourself from your WhatsApp conversations
here's how you can export your chat history: https://faq.whatsapp.com/1180414079177245/?cms_platform=andr...
advaith08··on Show HN: WhatsApp-Llama: A clone of yourself from your WhatsApp conversations
wow thats cool! Do you have the code put out somewhere?
advaith08··on Show HN: WhatsApp-Llama: A clone of yourself from your WhatsApp conversations
Yeah, this is definitely a dicey ethical question. Would be interested to know what guardrails you're considering for these digital avatars, and how you'll ensure that people use them in a healthy manner and dont get dependent on them.
advaith08··on Show HN: WhatsApp-Llama: A clone of yourself from your WhatsApp conversations
Thank you, Ill check it out!

Yeah, this can be extended to create a "simulation game" of us and our friends. This paper (Interactive Simulacra of Human Behaviour https://arxiv.org/abs/2304.03442 ) has a setup on how we could create a Sims game with us as the characters

advaith08··on Show HN: WhatsApp-Llama: A clone of yourself from your WhatsApp conversations
Yeah I wondered if few shot prompting would yield better results than finetuning. For the amount of finetuning I've done (1 epoch, 7B model with 4 bit quantization), I think it might be comparable. But if we scale this to a bigger model and longer training times, I think finetuning should produce much better results. Hoping someone with access to compute will try it out and update us!
advaith08··on Show HN: WhatsApp-Llama: A clone of yourself from your WhatsApp conversations
Hahaha a friend told me this but I havent watched SV. Will do so immediately
advaith08··on Show HN: WhatsApp-Llama: A clone of yourself from your WhatsApp conversations
oh yeah definitely. Do you know how I can get access to one for cheap though? I burnt through $150 just on this exercise with a P100 on GCP