Perhaps also symmetric "freedom to learn" from OpenAI models, with some provisions / naming convention? U.S. labs are limited in this way, while labs in China are not.
2,070 karma · joined November 17, 2011
... "What I can not create, I do not understand." ...
... "Know how to solve every problem that has been solved." ...
... "Four no's. Five clues." ...
... "Reality is that which, when you stop believing in it, doesn’t go away." ...
... "A human being is a part of the whole called by us universe, a part limited in time and space." ...
Perhaps also symmetric "freedom to learn" from OpenAI models, with some provisions / naming convention? U.S. labs are limited in this way, while labs in China are not.
- UMG Recordings, Inc.
- Capitol Records, LLC
- Concord Bicycle Assets, LLC
- CMGI Recorded Music Assets LLC
- Sony Music Entertainment
- Arista Music
Taken from: https://cdn.arstechnica.net/wp-content/uploads/2025/03/UMG-v...
And no, I don't think the knowledge of language is necessary. To give a concrete example, tokens from TinyStories dataset (the dataset size is ~1GB) are known to be sufficient to bootstrap basic language.
And note, the system is now directly competing with "interns". Once the accuracy is competitive (is it already?) with an average "intern", there'd be fewer reasons to hire paid "interns" (more expensive than $200/month). Which is maybe a good thing? Fewer kids wasting their time/eyes looking at the computer screens?
This is far from the case — many areas are characterized by heavy-tailed loss distributions, where extreme negative consequences could really ruin the day and erase any efficiency gains.
In your particular case the prompt would look something like: <pubmed dump> what are the plants that aren't poisonous to most people?
A general reasoner would recover language and relevant world model from pubmed dump. And then would proceed to reason about it, to perform the task.
It doesn't look like a particularly efficient process.
On the other hand, my take on it, the ability to do reasoning in a long context is a general capability. And my guess is that it can be bootstrapped from scratch, without having to do training on all of the internet or having to distill models trained on the internet.
Emergent tool use from multi-agent interaction is a good example - https://openai.com/index/emergent-tool-use/
> Thought about large prime check for 3m 52s: "Despite its interesting pattern of digits, 12,345,678,910,987,654,321 is definitely not prime. It is a large composite number with no small prime factors."
Feels like this Online Encyclopedia of Integer Sequences (OEIS) would be a good candidate for a hallucination benchmark...
This brings computers into the classroom, and once they’re available, it is a slippery slope. It is easier for teachers to have students use semi-gamified "educational" apps rather than engage themselves.
Example for K-2 - https://www.cde.ca.gov/be/st/ss/documents/csstandards.pdf:
K-2.CS.1 Select and operate computing devices that perform a variety of tasks accurately and quickly based on user needs and preferences.
K-2.CS.2 Explain the functions of common hardware and software components of computing systems.
K-2.CS.3 Describe basic hardware and software problems using accurate terminology.
K-2.NI.4 Model and describe how people connect to other people, places, information and ideas through a network.
...
K–2 K-2.AP.12 Create programs with sequences of commands and simple loops, to express ideas or address a problem
K-2.IC.20 Describe approaches and rationales for keeping login information private, and for logging off of devices appropriatelyImagine my frustration one day, when I've discovered that my kindergartner has full access to a brand-new, shiny iPad during class. Despite complaints from parents, the teacher refused to reduce iPad usage (or even activate Screen Distance and Screen Time controls on the iPad, or share usage statistics).
The only thing that I've learned, this is all in line with California’s state-approved computer literacy recommendations.
Like FTC, I estimate that banning these would save U.S. consumers millions of hours they currently spend searching and clicking on pointless coupons on their phones before making purchases. It would also increase happiness, as it's extremely annoying to pay $20 extra, knowing that a lower price is available if only you spent ten minutes struggling with a store's website on your phone.
Whoever invented this is evil and is destroying happiness.
Like FTC, I estimate that banning these would save U.S. consumers millions of hours they currently spend searching and clicking on pointless coupons on their phones before making purchases. It would also increase happiness, as it's extremely annoying to pay $20 extra, knowing that a lower price is available if only you spent ten minutes struggling with a store's website on your phone.
Whoever invented this is evil and is destroying happiness.
For something indoors, yes, I can see how low sampling frequency gets very limiting. And 192 microphones, that's really pushing it. Love it.
The $2/mic vs $0.5/mic argument is a fun one. You've obviously poured enormous amount of engineering in there, involving PCB design, FPGA and network programming, writing custom CUDA kernels, signal processing, PyTorch, the list goes on. And you've had 4090 plugged in your PC in 2023. Classic hobbit in a mithril vest ;)
I understand that ICS-52000 is a relatively low cost ($2/100pcs) and there are even breakout boards available with 4 microphones, which can be chained to 8 or 16, like https://www.cdiweb.com/datasheets/notwired/ds-nw-aud-ics5200...
Then you can take Jetson (or any I2S capable hardware with DSP or GPU on it) and chain 16 microphones per I2S port. It would seem a lot easier to assemble and probgam, if comared to FPGA setup.
And it is reasonable to have failure cases. But systems should fail gracefully. This wasn't a graceful failure.
If you'd like an insight of what roughly can happen inside this computation, a short story from Karpathy is not the worst read - http://karpathy.github.io/2021/03/27/forward-pass/
The income cap on getting the clean vehicle rebates is $135k ($200k joint filers). And I'm not sure about the federal rebates. Tesla doesn't offer 0% financing, current Cybertruck APR deal is reported to be 5.29% for up to 72 months. So I don't see how someone with the income under the rebate cutoff can afford that $100k car or the financing option. The delta between the number of rebates (Federal EV vs Clean Vehicle/CA) may allow to estimate, how many of these are corporate (pre-income tax + rebate?) purchases.
And these "Cox Automotive estimates", are these reliable numbers that had been confirmed by Tesla earnings, or it is a "best guess by influencers" type of information?
I'm curious, what is the alternative that you are considering? I've been delaying an upgrade to electric for some time. And now, a car manufacturer that is contributing to the making of another Jan 6th, 2021 is not an option, in my opinion.
We see similar situation in automotive. Other companies do allow to keep Tesla in check, so there's less opportunity to force "Cybertrucks" onto the market as the only option.
It's like taking an executable (.so module, firmware blob) and releasing it under permissive license, so anyone could disassemble, modify and hack it. And then disclosing what programming languages were used and pointing at a few libraries. And then saying that no, actual source code is not going to be released.