HNHacker News
TopNewBestAskShowJobs

rawrawrawrr

201 karma · joined February 13, 2022

submissionscomments
rawrawrawrr··on Alibaba's OpenAI Challenger: The New AI Reasoning Titan
Non-blogspam link: https://qwenlm.github.io/blog/qwq-32b-preview/

discussion: https://news.ycombinator.com/item?id=42259184

rawrawrawrr··on Zapata AI Ceases Operations
I wish I knew that there existed a public quantum AI company as an easy short before they went out of business.
rawrawrawrr··on Reflection 70B, the top open-source model
The model page says the prompt is:

The system prompt used for training this model is:

   You are a world-class AI system, capable of complex reasoning and reflection. Reason through the query inside <thinking> tags, and then provide your final response inside <output> tags. If you detect that you made a mistake in your reasoning at any point, correct yourself inside <reflection> tags.

from: https://huggingface.co/mattshumer/Reflection-70B
rawrawrawrr··on Reflection 70B, the top open-source model
Nice! It would be a better benchmark to compare this prompt (w/ gpt-4o, claude) with whatever the original model was compared to.
rawrawrawrr··on Apache Zeppelin
Google Colab has this, I wouldn't be surprised if there was a Jupyter widget to implement something similar.

Edit: looks like Mercury (A jupyter extension) has them: https://runmercury.com/docs/input-widgets/

rawrawrawrr··on Phind-405B and faster, high quality AI answers for everyone
Title says "for everyone", but post says "Phind-405B is available now for all Phind Pro users". I guess everyone on earth has paid for Phind :)
rawrawrawrr··on Ilya Sutskever's SSI Inc raises $1B
Are you a VC? If they really didn't care about their investment exits, that would be crazy.
rawrawrawrr··on SAM 2: Segment Anything in Images and Videos
Can run on CPU (slower) or AMD GPUs.
rawrawrawrr··on SAM 2: Segment Anything in Images and Videos
Might have issues if you're from Texas or Illinois due to their local laws.
rawrawrawrr··on ChatTTS-Best open source TTS Model
Video games, for one.
rawrawrawrr··on Sam Altman is showing us who he really is
Parody is covered under fair use.
rawrawrawrr··on HMT: Hierarchical Memory Transformer for Long Context Language Processing
That's kinda hilarious, because I think the book was exactly wrong in its predictions. This can be evidenced by the continuous failures of his AI company, Numenta.
rawrawrawrr··on The Ramen Lord
Ramen Lord is an awesome guy. The sheer devotion is something to marvel, as someone who is constantly distracted by Reddit/Hacker News/latest tech gizmo of the day.

Fun fact: in Japan, ramen is considered (Japanized) Chinese, but Japanese in the US. It's an interesting game of cultural telephone.

rawrawrawrr··on Science fiction authors were excluded from awards for fear of offending China
Babel is a decent sci-fi book, it's actually quite anti-capitalist. I'm not making a political statement.
rawrawrawrr··on Show HN: A platform for remote piano lessons based on the Web MIDI API
Did you write your own VSTI, or what did you use?
rawrawrawrr··on Show HN: A platform for remote piano lessons based on the Web MIDI API
It's not a big deal imo. I started on digital piano, moved over a few years to a real piano. The brain adapts quickly.
rawrawrawrr··on Sam Altman Seeks Trillions of Dollars to Reshape Business of Chips and AI
It depends. Right now once we hit 6-8 bit precision inference, H100s/A100s are not memory-bound, but compute-bound.
rawrawrawrr··on Sapling – A VCS from Meta
Git is to SVN as Sapling is to Git. It's really a great tool, makes version control a lot more productive imo.
rawrawrawrr··on One-Pedal Driving Explained
If you're planning on using the brake pedal, then that's not one pedal driving.
rawrawrawrr··on Beijing criticises Netherlands’ move to block ASML exports to China
> Chinese foreign ministry spokesman Wang Wenbin on Tuesday urged the Netherlands "to be impartial, respect market principles and the law, take practical actions to protect the common interests of both countries and their companies and maintain the stability of international supply chains".

This amuses me greatly. I had no idea China is explicitly pro-market capitalism.

rawrawrawrr··on Amazon Unveils Graviton4: A 96-Core ARM CPU with 536.7 GBps Memory Bandwidth
Fun fact, a Nvidia RTX 4090 has ~1TBps memory bandwidth, for an apples to oranges comparison.
rawrawrawrr··on MeshGPT: Generating triangle meshes with decoder-only transformers
It's research, not meant for commercialization. The main point is in the process, not necessarily the output.
rawrawrawrr··on Kotlin Coroutines vs. Threads Performance Benchmark
Doesn't seem like there's anything wrong with the methodology in the quote. It's perfectly fine to compare two things, even if one is not meant for the task.
rawrawrawrr··on Google to invest up to $2B in Anthropic
$6 billion dollars raised in 2 months for their series C is blowing my mind. What does Anthropic have that OpenAI or other LLM startups don't? What other companies have raised that much in a single round?
rawrawrawrr··on Show HN: Sheet Music Management App
I understand that this is an ad for airsequel, but as a musician and programmer I have no idea what this is. What kind of sheet music does it accept? Do I need to manually enter in all the fields?
rawrawrawrr··on Show HN: Lantern – a PostgreSQL vector database for building AI applications
This is fine, its a standard enterprise marketing technique. Watch a webinar, get a gift card.
rawrawrawrr··on Redesigned Google Fonts website
The old design made it easy to copy paste some CSS to use Google's CDN to add fonts to a website. With the new design, I have no idea how to do this. Am I stupid?

Edit: Pressing Select [font name] + adds it to a "shopping cart", where it displays the code to copy it to a website. I am dumb.

rawrawrawrr··on Slack AI
I found it quite useful for summarizing the last X hours of chat messages.
rawrawrawrr··on Phind (code beating GPT4) seems to have used WizardLM's finetuned checkpoint
To preempt this drama, Phind claims that they were inspired by WizardLM's technique of generating datasets using GPT-4, but didn't use their model or data.
rawrawrawrr··on Code Llama, a state-of-the-art large language model for coding
> Not a bad context

A little understated, this is state of the art. GPT-4 only offers 32k.

Page 1 of 2Next →