HNHacker News
TopNewBestAskShowJobs

tim_sw

12,831 karma · joined March 22, 2007

Yet another Tim
submissionscomments
tim_sw··on Former Google CEO Eric Schmidt's Leaked Stanford Talk
False. Eric Schmidt is one of the sharper operators in tech and his effect can be seen when Google started declining after 2011 when he stepped down.
tim_sw··on LangChain vs. LlamaIndex
My 2 cents - don’t rely on these frameworks and just do it yourself (or pick libraries like Instructor over these frameworks)

I think both have the wrong abstractions for people to build more complex workflows and use cases beyond demos.

tim_sw··on Tomas Mikolov on origin of word2vec and seq2seq
Pastebin https://pastebin.com/StEkp0YT
tim_sw··on Open source Datadog rival SigNoz lands on the cloud with $6.5M investment
How does this compare with HyperDX?

https://news.ycombinator.com/item?id=37558357

tim_sw··on TSA officers accused of stealing from passengers
This is more common than reported. I’ve lost a semi expensive present, and have heard anecdotal stories of people losing handbags, jewelry, higher end clothing/shoes.

It seems to be more prevalent in the US/TSA than other first world countries.

tim_sw··on Nvidia H100 GPUs: Supply and Demand
This is a very high quality writeup.
tim_sw··on Llama 2: an open-source LLM
this is not a duplicate - it's not the announcement post, it's a deep dive into the paper from the POV of Hugging Face's RL lead.
tim_sw··on Llama 2
Defense against tiktok and bytedance
tim_sw··on Moneyball for Software Teams: Quantifying Dev Performance
The blog’s name is “Software The Hard Way” and this post lives up to that name
tim_sw··on Databricks Strikes $1.3B Deal for Generative AI Startup MosaicML
The devil is in the details. Training large LLMs requires a lot of custom infra (handling GPUs going down, efficiently pushing data to keep the accelators busy, deciding on which mechanism of parallelizing model training is better - data vs model parallelism or both, tuning hyperparams of optimizers which can be different for larger batch sizes, etc)

Mosaic is one of the better providers for this. AWS is nowhere near ready at this current point in time, it is pretty much a "dumb" infra provider in large LLM training at this point. (Of course they won't be standing still and will prob acquire that capability one way or another)

tim_sw··on The Instant Pot Failed Because It Was a Good Product
https://archive.ph/bj54P
tim_sw··on Microsoft and OpenAI Forge Awkward Partnership as Tech’s New Power Couple
https://archive.ph/1NGwV
tim_sw··on Motorola ditched cell phones and found a lucrative second act
https://archive.md/xAean
tim_sw··on A prompt pattern catalog to enhance prompt engineering with ChatGPT
While somewhat useful - there are no systematic ablation and comparison studies, datasets, and quantitative evals here. It's all anecdotal, so should probably be a blogpost and not a "paper"
tim_sw··on Nvidia created the chip powering the generative AI boom
https://archive.md/GGVpH
tim_sw··on Evaluating Verifiability in Generative Search Engines
We find that responses from existing generative search engines are fluent and appear informative, but frequently contain unsupported statements and inaccurate citations: on average, a mere 51.5% of generated sentences are fully supported by citations and only 74.5% of citations support their associated sentence. We believe that these results are concerningly low for systems that may serve as a primary tool for information-seeking users, especially given their facade of trustworthiness. We hope that our results further motivate the development of trustworthy generative search engines and help researchers and users better understand the shortcomings of existing commercial systems.
tim_sw··on Amazon Overhauls Delivery Network – From National to Regional
https://archive.ph/4YOKT
tim_sw··on Google Launches AI Supercomputer Powered by Nvidia H100 GPUs
this looks like it's for GCP. TPUs are used for most internal workloads. It's available externally but some of the papercuts and devex without the TPU/TF team helping you can be more painful than using Nvidia/CUDA
tim_sw··on Codename Burnham: Amazon has a secret new home robot with ChatGPT-like features
https://archive.ph/AJQzE#selection-1797.6-1800.0
tim_sw··on American children are drowning in self-esteem (2016)
https://archive.md/FUNLn
tim_sw··on Microsoft agrees to stop bundling Teams with Office
https://archive.ph/VYpFZ
tim_sw··on Google’s maxtext – A simple, performant and scalable Jax LLM
Pretty high model flop utilization:

—————

MaxText is a high performance, arbitrarily scalable, open-source, simple, easily forkable, well-tested, batteries included LLM written in pure Python/Jax and targeting Google Cloud TPUs. MaxText typically achieves 55% to 60% model-flop utilization and scales from single host to very large clusters while staying simple and "optimization-free" thanks to the power of Jax and the XLA compiler.

tim_sw··on LangChain taps Sequoia to lead funding round at a valuation of $200M
https://archive.ph/PNRaO
tim_sw··on China hit by surge in Belt and Road bad loans
https://archive.ph/MzEq1
tim_sw··on AWS facing short-term headwinds as companies are cautious on spending
https://archive.md/03XXL#selection-237.3-237.90
tim_sw··on The Lure of Singapore: Chinese Flock to ‘Asia’s Switzerland’
https://archive.md/EA0CL
tim_sw··on Elon Musk Focusing Twitter on AI
https://archive.ph/tA46W
tim_sw··on The Kissimmee River has been brought back to life, and wildlife is thriving
https://archive.md/KGOUQ
tim_sw··on Faiss: A library for efficient similarity search
How are people using vector DBs in production? Do you typically use and manage Faiss indexes alone or use something like Milvus, Pinecone, Weaviate, or Chroma?
tim_sw··on Replit and Google Cloud Partner to Advance Generative AI for SW Dev
Fun tweet from @amasad

https://twitter.com/amasad/status/1640831001256689664?s=46&t...

Page 1 of 4Next →