Anyone else witnessing a panic inside NLP orgs of big tech companies?
old.reddit.com
old.reddit.com
1. You don't need to take everyone's job. You just need to take a shitload of people's jobs. I think a lot of our current sociological problems, problems associated with wealth inequality, etc., are due to the fact that lots of people no longer have competitive enough skills because technology made them obsolete.
2. The state of AI progress makes it impossible for humans in many fields to keep up. Imagine if you spent your entire career working on NLP, and now find GPT-4 will run rings around whatever you've done. What do you do now?
I mean, does anyone think that things like human translators, medical transcriptionists, court reporters, etc. will exist as jobs at all in 10-20 years? Maybe 1-2 years? It's fine to say "great, that can free up people for other thing", but given our current economic systems, how are these people supposed to eat?
EDIT: I see a lot of responses along the lines of "Have you seen the bugs Google/Bing Translate has?" or "Imagine how frustrated you get with automated chat bots now!" Gang, the whole point is that GPT-4 blows these existing models out of the water. People who work in these fields are blown away by the huge advances in quality of output in just a short time. So I'm a bit baffled why folks are comparing the annoyances of ordering at a McDonald's automated kiosk to what state-of-the-art LLMs can do. And reminder that the first LLM was only created in 2018.
Tangent, I’m looking for some sci fi about this topic. Any suggestions?
But we can technically do much better waterproof concrete or whatever - however our incentives are also not aligned in the same ways.
Here's a tangential link to monks building a Gothic cathedral with modern machines: https://carmelitegothic.com/
97% of jobs used to be working on the farm. Now it's something like 2%.
Ted Faro is alive today. We just aren't sure which of the FRANK'S he's at.
Ted Faro was a horrible human being blinded by delusions of grandeur, but he wasn’t “evil” - he was even convinced he was saving humanity by ending the threat of both war and climate change.
I don’t see GPT itself as representing a new Faro Plague, but I do see a lot of wannabe Ted Faros making the decisions at the top.
If LLMs come even close to achieving their short-term potential, we’re unleashing a bigger destabilizing force on the world than the smartphone/social media combo - and the world of 202x seems blatantly incapable of absorbing that level of disruption.
That is no longer the case.
Think of it instead as cognitive habitat. Sure, there has been habitat loss in the past, but those losses have been offset by habitat gains elsewhere.
This time, I don't see anywhere for habitat gains to come, and I see a massive, enormous, looming (ha!) cognitive habitat loss.
-- EDIT:
Reply to reply, posted as edit because I hit the HN rate limit:
> Your job didn't exist then. Mine didn't, either.
Yes, that was my point. New habitat opened up. I infer (but cannot prove) that the same will not be true this time. At the least, the newly-created habitat (prompt engineer, etc.) will be miniscule compared to what has been lost.
Reasoning from historical lessons learned during the introduction of TNT was of course tried when nuclear arms were created as well. Yet lessons from the TNT era proved ineffective at describing the world that was ushered into being. Firebombing, while as destructive as a small nuclear warhead, was hard, requiring fantastic air and ground support to achieve. Whereas dropping nukes is easy. It was precisely that ease-of-use that raised the profile of game theory and Mutually Assured Destruction, tit-for-tat, and all the other novelties occurrent in the nuclear world and not the one it supplanted.
Arguing from what happened with looms feels like the sort of undergrad maneuver that makes for a good term paper, but lousy economic policy. So many disanalogies.
Your job didn't exist then. Mine didn't, either.
This prediction has occurred with every technology revolution. It hasn't been borne out yet.
> the sort of undergrad maneuver
It's a variation of the broken window fallacy.
So what? You are performing 'induction from history', which is possibly the hand-waviest possible means of estimating what is next to occur.
Discontinuities occur. Fire gets tamed. Alphabets get invented. What went before is only a solid guide to the future absent any major disruption to the status quo. There is no a-priori reason to think that this time will be the same, either. Burden of proof is yours.
> It's a variation of the broken window fallacy.
I appreciate parsimony as much as the next academic but I'd appreciate you fleshing out your position here, so I can take it apart at the joints, in the custom and manner of my people >:)
The "this time it's gonna be completely unlike anything seen before!" is much more suspect.
> fleshing out your position here
https://en.wikipedia.org/wiki/Parable_of_the_broken_window
The parallels with your theory are unmistakable. Prosperity doesn't come from jobs that are little more than make-work.
I propose that you invest in a more convincing line of argument. The burden of proof lies heavy upon you.
Secondly -- for the life of me -- I don't see how we got from "prosperity doesn't come from jobs that are little more than make-work" (a claim, by the way, that Keynes would take exception to) to the view that automating most of the intellectual work on the planet will have nugatory impact, or that we'll all just vy to become celebrated Twitch streamers or influencers or whatever (assuming that synthetic influencers don't take off -- oh, wait, they did: https://www.synthesia.io/glossary/ai-influencer)
Even were you correct (and Keynes wrong), the instantaneous conversion of meaningful labour --- journalism, counselling, engineering -- into, as you say, "make-work" (the position I infer you are taking) would have tremendous cost.
Need I invoke Rawls here on widely-distributed self-respect/self-esteem? It's actually one of the prerequisites, he notes, for a stable civilization; https://academic.oup.com/book/32571/chapter-abstract/2703665...
At minimum, the psychological impact of such a transition would make the developed world's COVID hangover look like a day at the zoo.
Finally, the Parable of the Broken Window specifically refers to destructive work. Non-productive work is not covered. https://finshots.in/archive/dig-holes-and-get-paid-to-fill-t.... And that is to say nothing about how economic fruits are distributed -- a whole other matter, upon which, I again infer, you have no further comment.
Up until Louis Pasteur invented the germ theory of illness, it was broadly understood, across many different cultures, that disease had its origins in one or all of: witchcraft, possession, loose morals, blocked meridians, etc.
Were you to do historical induction on the spawning of illness theory, you might well conclude that no theory of illness would be scientifically verifiable. You might have argued that anyone claiming a radical change in medicine was deluded, alarmist, or simply excitable.
And you would have missed out on the multiple decades of extra health that you've had on account of antibiotics, sterile procedure, and disinfectant. Your induction from history would have caused you to miss the disanalogy.
Something to think about next time you're at the doctor.
As to distribution of economic fruits, as I mentioned before, replacing labor with machines made the US the most prosperous country in history, along with the richest poor people.
Your country, the USA, is very close to system collapse because of its inequal distribution of fruits. I daresay it makes good popcorn-time.
Sorry, but breaking a window and then fixing it is non-productive.
> Your country, the USA, is very close to system collapse because of its inequal distribution of fruits.
Hardly. If the US will collapse, it's because of the current leftist swing of the government engaging in ever-increasing wealth redistribution.
> I daresay it makes good popcorn-time.
It's the equal distribution of income countries that repeatedly collapse. France is in the news currently because they've discovered that the math of redistribution does not work, and the people who cannot accept the math are rioting.
Good night and good luck
And yea you can say that birds are dinosaurs but the transition was very messy.
I wonder how long it will take politics to appreciate the deep societal problems that may soon arise. Probably it will be too little, too late.
The luddites were trying to defend their livelihoods, communities etc. It's the same rational thing people are looking at now: How to survive
Fast forward to today. I received a solicitation for a charity in the mail the other day. They enclosed pictures of the poor kids. The kids were all wearing fashionable, spotless clothes in perfect condition.
That's what industrial textile machines gave us.
Any source for those starvation deaths? I would like to learn more about what prevented them from simply doing what the survivors did.
You can not easily buy a weaving machine (there are some second hand ones) or easily go to your local maker space and use the weaving machine to create the design you desire. Open source in the space of textile making is in its infancy even though there are some projects. I bet it is easier to get a low volume tape-out of some custom chips than it is to get a custom roll of textile. (you can get printing but that's not the same thing)
Textiles have become way cheaper and both higher quality (when demanded) and lower quality (cost saving fast fashion) and available in far higher quantities.
Some things were lost in this transition.
I've seen people use them to make rugs and things for sale in gift shops.
There's nothing stopping you from doing this.
My point was that access to the technological advancement has not trickled down and that this creates an imbalance of power.
Compare how easy it is to get a custom textile (not a custom print) made to how easy it is to get a custom PCB made (it is reasonably easy to etch a double sided board and multi-layer and flexible ones can easily be ordered online). The situation with regards to knitting is somewhat better.
Saying that a "a basic loom isn't hard to make" in a world of high speed air jet digital looms is equivalent to saying that perf-board still exists in a world of SMD components.
Honestly as a human who grasps how the economy works, this doesn't sound like a good thing, but I don't see any path to trying the fundamental changes that would be required for really good general AI to not be an absolute Depression generator.
The only thing I'm wondering is, will the wealthiest ones, who actually have any power to influence these fundamental thing, figure this out before it's too late? I really doubt your Musks and Bezoses would enjoy living out their lives on ring-fenced compounds or remote islands while the rest of the world devolves into the Hunger Games.
This is complete nonsense. An LLM has no ability to innovate or 'try new things'.
Just what the shit do you think Google/Microsoft/Amazon are pouring billions of dollars into machine learning/AI for? The first one that creates a self improving-self learning machine wins the game (or destroys the earth with a paperclip maximizer).
You and your human intelligence are not magic. You're biological hardware and software that a lot of people are spending a lot of time and effort on reproducing in a digital format.
This is (materialistic) nihilism. Materialism is a philosophy and not a new one, but an exceedingly sterile one, in my view. If you want to take that philosophical position you can, but others are free to reject it (and most people do) because it is only a philosophical position, not a proved description of reality (how can it be?).
Which really makes you wonder if F/OSS was a good idea in the first place.
I see where you’re coming from, but is this really the main source of the inequality?
Based on numbers relating to workers’ diminishing share of profits, it seems to be that the capital class has been able to take a bigger piece of the profit pie without sharing. In the past, companies have shared profits more widely due to benevolence (it happens), government edict (e.g., ww2 era), or social/political pressure (e.g., post-war boom).
Fwiw, I think that the mid-20th century build up of the middle class was an anomaly (sadly), and perhaps we are just reverting to the norm in terms of capital class and worker class extremes.
I see tons of super skilled folks still getting financially fucked by the capital class simply because there is no real option other than to try to attempt to become part of the capital class.
Yes, more of this money is going, instead of middle-class workers, straight to the capital class who own the "machines" that do the work people used to do. Except instead of it being a factory that makes industrial machines owned by some wealthy industrialist, the machines are things like Google and AWS and the owners are the small number of people with significant stock holdings.
It's really striking though that a person graduating high school in say, 1970, could easily pick from a number of career choices even without doing college or even learning an in-demand trade, like plumbing, welding, etc. Factory work still existed and had a natural career progression that wasn't basically minimum wage, and the same went for retail. Sure, McDonalds burger flippers didn't expect then to own the restaurant in 10 years, but you could take lots of retail or clerical jobs, advance through hard work and support a family on those wages. Those are the days that are super gone and I totally agree with you both that something has changed for the worse for everyone who's not already wealthy.
Only in certain places, and only mostly due to crazy policies that made housing ridiculously unaffordable. I'm in an area where my barber lives on 10 acres of land he didn't inherit and together with his wife raises two children. This type of relaxed life is possible to do in wide swathes of the country outside of the tier-one cities that have global competition trying to get in and live there, as long as you make prudent choices.
I think 20- to 30-something engineers who have spent their entire adult lives in major coastal cities have a huge blind spot to how middle America lives.
[0]: https://dqydj.com/average-median-top-household-income-percen...
In general, I'd argue minimum wage jobs aren't anymore a stepping stone to some sustainable good job; they're basically viewed like consumables by companies. Even someone with a decade of experience, say, at several retail stores or restaurants, can't expect to be offered a position making $50,000 a year plus benefits just for having done his job well every day. By contrast, a dedicated factory worker with a decade of experience 50 years ago could expect to have advanced somewhat, and would expect continued advancement. Today everyone working in retail, restaurants, etc. all know that if they're going to do any better it's going to be by leaving that sector, via learning a trade, going to college, or perhaps founding their own small business. All things which were good options in the past too, but advancement was once a realistic expectation too.
The only situation I can imagine would be if the wife has a government job with extremely generous health insurance subsidies.
There’s always another appointment lined up before and after mine, so I guess he’s pulling 6 figures without much sweat.
I live in the principal city of the local Metropolitan Statistical Area; it's by no means a big city, but it's representative of many small cities around the country. My barber lives out in the county, outside of city limits, where it is much more rural and one can indeed buy 10 acres for not a whole lot of money.
I believe his wife is a schoolteacher; I don't believe public employee benefits are especially generous in this state.
Ask them what their deductible/oop max, and how they get that insurance, and I bet you will have your answer for how your friend can afford to raise a family of 4 as a barber and buy and live on 10 acres of land. I doubt a 2 barber couple could pull it off. The security/benefits of one half of a couple being a government employee is pretty valuable.
Bet you can get the same for less than $20 within 3 miles (and possibly far less).
I paid $10 for my last haircut ($6 two months ago - thank you, Joe Biden). Luckily there's another shop nearby that still charges $5.
I’m not just paying for a haircut, it’s an experience that’s worth every penny.
This isn't specific to cities.
In fact, people in rural areas are worse impacted, because the rise of Walmart, Dollar General, and others funnel money out of their towns that would have otherwise enable many local families to capture the profits from local spending. Today a lot of that spending goes mostly to those companies, and only a fraction of the money stays, in the form of a few low-wage jobs.
I'm not saying it's impossible to not live in poverty. I'm just saying it's much much harder, because "advancement" is obsolete in a lot of occupations where it used to be a thing.
Isn't this rather a strong argument for the claim that what high school as of today teaches is a strong mismatch with what the labour market demands? In other words: the pupils are taught skills for many years of their life that are rather worthless for the job market.
But high school and college both got dumbed down, and now an education at a state university is comparable to high school in the first half of the twentieth century.
Why won't tradesman cut their prices?
School teaches everybody to read, write and reason about things in general to a decent level. You can't teach high school kids the basics of all the careers out there beyond stuff that's generally applicable - you wouldn't have the time or the equipment. And schools do often have elective shop, cooking, electronics etc classes for those who want to do them.
Ha, those were cut from school before I was in high school decades ago.
I don't think UBI is the solution. Nor is squeezing people more than they are being squeezed. Efficiency and productivity are good things. What is wasteful are things like make-work programs like the DMV or other govt office. That crap needs to be automated away. Hospitals need more funding. Schools are unclear. I think schools would benefit from privatization. I don't think the same of hospitals. Not sure why.
And outside of electronics a lot of physical-goods/land stuff is less attainable in many places in the country anyway.
The problem with non-redistributive approaches is that generational wealth rarely goes away. So if you don't have it, the number of people who don't have to try to out-spend you for whatever you want only goes up as time passes.
Today, that's close to impossible unless you take student loans, go through the gauntlet of getting a higher paying job, and then have to grapple with home prices assuming you don't live in a place with reasonably affordable home ownership.
If you think the DMV and government offices are "make-work" then we need to start with cutting military spending because it's the biggest "make-work" government program we have that the vast majority of U.S. spending goes to.
[0] https://www.thebalancemoney.com/u-s-federal-budget-breakdown...
We can't change mandatory spending or we have millions of elderly people dying because social security vanishes into thin air.
Some people call it a Ponzi scheme, but it’s a pool, as designed.
The military is a big jobs program and all the industrial support generates a lot of economic activity. Better if they just dropped it in the ocean, though. Too much temptation to test it.
I'm pretty sure the latter 3 are more important to standard of living than the former.
Ergo, I believe your claim that standard of living is superior to 1970 is false. Having a shiny iPhone to distract you from the fact that you're homeless, sick and starving is not a step up.
[0] https://mpost.io/wp-content/uploads/975d2151-7def-45a4-8c52-...
Well, we (in the US), have slowly been privatizing them and it’s bad! If you look at test results though, it’s great! Because private charter schools can drop underperforming students before the end of the term and artificially inflate their numbers. There are many more reasons that education with a profit motive isn’t better than without. I suggest maybe reading up on this before casually suggesting how you think we should radically erode our institution.
Consider the elephant in the room:
https://www.federalbudgetinpictures.com/federal-spending-per...
Where does that money come from?
The government spent it.
In the current world, where do you think a lot of the capital class is able to get their capital?
Technological progress, and especially the Internet, has made much bigger markets out of what were previously lots of little markets, and now th "winner take all/most" dynamics make it so that where you previously could have, for example, lots of "winners" in every city (e.g. local newspapers selling classified ads), where now Google, FB and Amazon gobble up most ad dollars - I think someone posted that Amazon's ad business alone is bigger than all US (maybe more than that?) newspaper ad businesses.
(Yes, they were printing bibles; radio and TV and the internet originally were going to be educational ...)
I submit that it is not hyperbole to say that probably 95% of all global human problems can have their root cause traced to poverty. That is not a scientific number, so don't ask for a citation (it ain't happening).
Anyone who only has the skills to affect the lives and/or environments of the people in their immediate surrounding are going to find themselves on the 'have nots' end of the spectrum in the coming decades.
But looking back further - ohhhhhh yeah dude. Oh yeah. Totally. For a brief period my mom worked at a BofA facility that _processed paper checks_. Like they had a whole big office for it. That's completely 100% gone now. The checks get scanned at the point of entry (cash registers, teller counters, etc) and then shredded.
https://en.wikipedia.org/wiki/Stabilization_Act_of_1942
In short the Federal government broke the individual laborers and labor unions to make sure they could afford to win World War 2.
Yes this tech may not eliminate jobs, but the jobs remaining will be indefinite less skilled easier to replace, or outsource
A few counter-notes
- Google translate and its ilk have already significantly cut down the number of translators required for multinational companies. Google translate in 2006 is also a bad example, it really only got excellent in the past few years.
- I would trust GPT to write the first draft, and then hire a translator to check it. That goes from many billable hours to one, or two. That is a material loss of work for said translator.
- High profile translations, as your example is, are a sharp minority of existing translator jobs.
There will always be a need for careful human translation work in such things as legal documents for government work but those positions will become even more competitive.
Recently I needed a couple payslips translated from German to English and then certified. Google Translate can translate, but cannot certify.
Professional licensed translator took significant amount of money and sent me a document for review with most numbers somehow mixed up. It took days of back and forth over the email for them to fix it.
And it doesn't mean that the replacements will be much better, or even as good as the Humana they replace. They will probably suck in ways that will become familiar and predictable, and at the same time irritating and inescapable. Think of the outsourced, automated voice systems at your doctor's office, self-checkout at the grocery store, those touchscreen kiosks at McDonalds, etc.
I already find myself wanting to scream
> GIVE ME A FUCKING HUMAN BEING
every now and then. That's only going to get worse.
not only am I sure people will have no trouble eating, I'm willing to wager obesity goes up
i'm not teasing, I don't think your worry is warranted
Before mechanical alarm clocks, there were people paid to tap on windows to wake them up.
As for human translators, the need for them far, far exceeds the number of them. Have you ever needed translation help? I sure have, but no human translator was available or was too expensive.
This is probably the real problem. Translators are payed shit nowadays for what is a really high-skill job. I have translators in the extended family who had to give up on that line of work because the pay wouldn’t sustain them anymore.
the issue is that many translation jobs will, and already are, being replaced with 'proofread machine translation output' jobs that simply don't pay enough. translation checking is careful, detailed work that often takes almost as much time as translating passages yourself, yet it pays a third or less of the rate because 'the machine is doing most of the work.'
An adequate AI translation is a lot better than no translation.
e.g. Italian citizenship can cost as much as a brand new car in Brazil and almost half of that cost could come from certified translation hurdles.
I somehow imagine it'll be the worst of both worlds but I'm a glass half empty kind of guy.
Everybody who doesn't have to do physical labour - including me - should be happy and grateful for that privilege - not try to rob even more from physical labourers, who in the end, create everything.
I have been doing NLP since 1993. Before ca. 1996, there were mostly rule-based systems that were just toys. They lacked robustness. Then statistical systems came up and things like spell-checking (considering context when doing it), part of speech tagging and eventually even parsing started to work. Back then, people could still only analyze sentences with fewer than 40 words - the rest was often cut off. Then came more and more advanced machine learning models (decision trees, HMMs, CRFs), first a whole zoo, and then support vector regressors (SVM/SVR) ate everything else for breakfast. Then in machine learning a revival of neural networks happened, because better training algorithms were discovered, more data became available and cheap GPUs were suddenly available because kids needed them for computer games. This led to what some call the ¨deep learning revolution¨. Tasks like speech recognition where people for decades tried to squeeze out another half percent drop in error rate suddenly made huge jumps, improving quality by 35% - so jaws dropped. (But today's models like BERT still only can process 512 words of text.)
So it is understandable that people worry at several ends. To lose jobs, to render ¨NLP redundant¨. I think that is not merited. Deep neural models have their own set of problems, which need to be solved. In particular, lack of transparency and presence of different types of bias, but also the size and energy consumption. Another issue is that for many tasks, no much data is actually available. The big corps like Google/Meta etc. push the big ¨foundational¨ models because in the consumer space there is ample data available. But there are very important segments (notably in the professional space - applications for accountants, lawyers, journalists, pharmacologists - all of which I have conducted projects in/for), where training data can be constructed for a lot of money, but it will never reach the size of the set of today`s FB likes. There will always be a need for people who build bespoke systems or customize systems for particular use cases or languages, so my bet is things will stay fun and exciting.
Also note that "NLP" is a vast field that includes much more than just word based language models. The field of propositional (logical) semantics, which is currently disconnected from the so-called foundational models, is much more fascinating than, say, chatGPT if you ask me. The people there, linguist-logicians like Johan Bos identify laws that restrict what a sentence can mean, given its structure, and rules how to map from sentences like "The man gave the girl a rose" to their functor-argument structure - something like "give(man_0, rose_1)¨ - which models the "who did what to whom?". When such symbolic approaches are integrated with neural foundational models, there will be a much bigger breakthrough than what we are seeing today (mark my words!). Because these tools, for instance Lambda Discourse Representation Theory and friends, permit you to represent how the meaning of "man bites dog" is different from "dog bites man".
So whereas today`s models SEEM a bit intelligent, but are actually only sophisticated statistical parrots, the future will bring something more principled. Then the ¨ "hallucinations" of models will stop.
I am glad I am in the field of NLP - it has been getting more exciting every year since 1993, and the best time still lies ahead!
This is a really important point. GPT-x knows nothing about my database schema, let alone the data in that schema, it can’t it learn it, and it’s too big to fit in a prompt.
Until we have AI that can learn on the job it’s like some delusional consultant who thinks they have all the solutions on day 1 and understands nothing about the business.
These general models can be fine tuned with domain specific data with a very small number of samples, and have surprisingly good transfer performance (beating classical models). New research like LORA/PEFT are making things like continuous finetuning possible. Statistical models also do a much better job at translating sentences to formal structure than the old ways ever did – so I wouldn't necessarily view those fields are disconnected.
I agree with the general sentiment, there are still major issues with the newer generation of models and things aren't fully cracked yet. But the scaling laws are saying there's still a lot of upside, even without new paradigms or architectural improvements.
If that’s a statistical parrot, well so are the rest of us.
But I think that is the issue...going forward wouldn't you hire a machine learning specialist rather than an NLP specialist for those problems? As far industry goes, is there any value in all the syntax/semantics/phonology theory NLP folks command?
Will linguistic knowledge not be needed at all? I don't want to speculate about the far-out future, but what I can safely say from industry experience that at any stage (1996 - now) there was always some extra gain to be had on top of the ((statistical | neural) - only) approach of the day by engineering hybrid solutions that at some level also exploit human-injected linguistic knowledge and human-injected business rules.
Next month at ECIR 2023 in Dublin, I will present a "shoot out" between a BERT model especially pre-trained (for months) and fine-tuned for document summarization of financial meetings (earnings calls) and a one-line POSIX shell script (two cascaded grep commands) written in 3 minutes by yours truly that also extrcts a summary - with surprising results...
You definitely see this in the weather space. Despite flashy headlines, AI has really failed to make much of a difference at core weather forecasting, because the specialized statistical systems that combine many numerical weather prediction models are so greatly refined to the generic forecasting problem that there is little room for improvement. And AI practitioners rarely even focus on the actual interesting problems in the field where we suspect there can be huge gains - like convective initiation (predicting where exactly storms will form and their potential phase trajectory, e.g. what is the probably it will go tornadic or produce large hail?). The reality is that meteorologists can refine the prediction task so precisely that you don't need innovative, brand new model architectures. And the crazy brand new pure DL/data-driven models like NVIDIA's FourCastNet or DeepMind's GraphCast have a long way to go to be a practical competitor to traditional NWP and basic post-processing/statistical bias correction.
* Verbal translation, where accuracy is usually important enough to want to also have a human onboard since humans still have an easier time with certain social clues.
* High-culture translation, where there's a lot to personal choice and explaining it. GPT can give out many versions but can't yet sufficiently explain its reasoning, nor would its tastes necessarily match that of humans.
* Technical translations for manuals and such. This market will be under severe threat from GPTs, though for high-accuracy cases one would still want a human editor just in case.
All in all, GPT will contract the market, but many human translators will be fine. There's still areas where you'd still want a human, and deskilling isn't a bug threat - a human can decide to immerse and get experience directly, and many will still do so by necessity.
In some cases, the quality may not go down or even go up (we've probably all seen some pretty bad human translations).
"AI" is not the only automation that threatens translation jobs: translation memories (plain ol' databases that remember past translations) have killed a lot of business for human translators, namely re-translating new versions of products, where not a lot has changed. Nowadays, they only get paid for the sentences that were modified compared to the previous version ("the diff"). SDL Trados is an example of this simple approach that is extremely effective, and used heavily e.g. by the European Commission's translation service.
I've said it before and I'll say it again. This right here is the crux of the issue. The only way people get to eat is if we change the economic systems.
Capitalism supercharged by AI will lead to misery for almost everyone, with a few Musks, Bezoses and Thiels being our neofeudal overlords.
The only hope is a complete break in economic systems, towards a techno-utopian socialism. AI could free us from having to do work to survive and usher in a Star Trek-like vision of the future where people are free to pursue their passions for their own sake.
We're at a fork in the road. We need to make sure we take the right path.
Things will get worse and worse until they boil over.
What can possibly be the benefit of requiring this constraint?
Remove the idea that this is necessary and watch how much relaxation comes to the deliberation on this topic.
"Current economic systems" will simply have to yield. Along with states. This has been obvious for decades now. Deep breaths, everybody. :-)
It's not "requiring this constraint". If you have some plausible pathway to get from our current system to some "Star Trek-like nirvana", I'm all ears. Hand-wavy-ness doesn't cut it.
> "Current economic systems" will simply have to yield.
Why? For most of human history there were a few overloads and everyone else was starving half the time. Even look at now. I'm guessing you probably live a decent existence in a decent country, but meanwhile billions of people around the world (who can't compete skills-wise with upper income countries) barely eke out an existence.
For the world that just lived through the pandemic, do you honestly see systems changing when worldwide cooperation and benevolence is a prerequisite?
My gut feeling is that AI is the 'social historic change' that will make UBI politically viable and a reality.
The problem is that we have created an economy where that is a bad thing.
The government of a wealthy country should ensure that its citizens are able to eat, and have a sheltered place to sleep, without them needing to work. Because the way things are going, there won't be enough work to go around. Even now, with the supposed "labour shortage" there are record numbers of homeless people, and people living paycheck-to-paycheck. Housing is more unaffordable than ever. Minimum wage is not keeping up with the economic realities.
Governments need to step in; they need to change policy so big corps are paying more taxes, and that tax goes to a basic income that can cover the cost of housing and the cost of food. Maybe not right away, maybe it starts at $100/month. But eventually the goal should be to get everyone on a basic income that can cover the necessities, then if they want to be able to enjoy luxuries (concerts, gourmet food, hobbies, streaming services, etc.) they can choose to work.
Computational irreducibility combined with human insatiability combined with ethics entails: a lot of work to be done. Everyone needs to get to work, we aren’t done yet. If robots could do all the work… well, god bless. There is SO much work to do. That makes me think the issue is not “robots taking jobs” but the focus and intention of the collective.
Any new hire is WAY more valuable now that they can use chatGPT. So why aren’t we hiring more people?
Well, one small issue: notice how you get taxed for hiring people a lot more than buying tech? That’s probably something to be fixed.
I don't really think of filing paperwork, or writing code, or meetings in general. It's the stuff that requires arms and legs and fingers that's our bottleneck. ChatGPT can't help with that. Robots can, but RL is lagging.
All that takes knowledge work. Physical too, but most of the cost today is knowledge work.
Now, my average tech staff can actually solve challenging coding problems that before could only be done by the few talented ones.
It does not solve all shortcomings, but it definitely shifts the collective cognitive load in the company.
Many of these jobs are cheap and easy to understand and quick to train in. These aren’t the kind of jobs people probably wanted, but they’ll be there.
Yes, they will be all called 'ai data labellers'.
For a long time, "People don't just want jobs, they want good jobs" was the slogan of industries that automated the boring stuff. Now AI is suddenly good at all the jobs people actually want and the only thing it can't do is self-improve. In an AI future, mediocre anything will not exist anymore.
Either you are brilliant enough to be sampling from 'out of distribution', or you're in the other 99 percent normies that follow the standard : "learn -> imitate -> internalize -> practice" cycle. That other 99% is now and eternally inferior to an AI.
Right! Aren't we all mediocre before we're excellent? Isn't every entry level job some version of trying to get past being mediocre? i.e. Isn't a jr developer "mediocre" compared to a senior dev? If AI replaces the jr dev, how will anyone become a senior dev if they never got the chance to gain experience to become less mediocre?
I used to do a job that was eventually automated. We did the one and only thing the computer couldn’t do - again and again in a very mechanical fashion.
It was a shit job. You might get promoted to supervisor - but that was like being a supervisor at McDonalds.
Why not treat the job seriously? Why didn’t the company use it as a way to recruit talent? Why didn’t the workers unionize?
Because we all knew it would be automated anyway.
We were treated like robots, and we treated the org like it was run by robots.
There’s a huge shadow over the economy that treats most new jobs like shit jobs.
Maybe the very, very basic transcription/translation stuff might go away, but arguably this race to the bottom market was already being killed by google translate as bad as it is anyway.
In areas where quality is required (eg. localizing video games from japanese to english and vis versa) people would be (justifiably) fussy about poor localization quality even when the translation was being done by humans, so I have to imagine that people will continue to be fussy and there will still be significant demand for quality job done by people who aren't just straight translating text, but localizing text for a different audience from another culture.
The way it works, someone would have to produce more original training data in the first place. In long term, it is AI that has to keep up, not the other way around.
> wealth inequality
Microsoft, OpenAI run closed for-profit LLMs that are inherently only possible thanks to creative work of all the people that might stand to lose jobs now. Not only it should be clear where the driving force for rising wealth inequality is going to be going forward—these companies’ effectively scraping original works of living humans and repackaging them for profit should be in violation of intellectual property law, if it isn’t already. Perhaps more people should start adding two and two together.
And if you think that AI is going to stay at LLMs and not eventually evolve into self learning/self modifying systems you are shortsighted.
On the contrary, the confusion is all yours. This is literally the reason plenty of photographers, illustrators were able to make money, as just one example. Without this any major publication could just grab whatever photo or artwork they saw fit, but they didn’t precisely because there’s such thing as IP law.
Before you argue for some form of communism without any intellectual property, ask yourself why people produced creative work in the first place (you know, all those works thanks to which an LLM can do its thing). Could it be because the fruit of their work was considered their intellectual property? As in, they were paid for it and were in control of it?
Now as soon as LLMs are trained on that work and can suddenly can produce derivatives cheaper, techbros are suddenly all like “let’s pretend IP law is some unfair thing that only benefitted the rich”. Those publishers now pay OpenAI/Microsoft for those very photographs and illustrations, that were taken for free, while the very people who created them would be losing jobs and gigs. Similar to book authors and other creative industries.
Do you really think this works in favor of decreasing the wealth gap? Not having to pay original creators for their work and instead paying a fraction of a penny to Microsoft while those creators starve? Aren’t you living in perpetual cognitive dissonance from these mental gymnastics?
No surprise non-tech people and especially creatives hate tech people more and more.
Or you could still have a human in the loop, for example a CTO reviewing their AI developer’s code. But now you can pay one person instead of 10.
That said, these tools still need some domain knowledge. It’s not at the point yet where anyone off the street can use it to accomplish an unfamiliar task.
There’s an old story about someone charging $1,000 to fix a machine. They did it in 10 seconds and the client complained. The consultant said “I get paid for knowing how to fix it”.
The same might be true for knowing how to prompt AI.
Yeah, exactly this. The funny thing about economic upheavals and industrial revolutions is that while people might get reallocated eventually (say, 30 years), that doesn't provide any comfort to the people who are getting upheaved now.
Somewhere toward that goal we became overly fixated with money/wealth and the pursuit of endless profits. Meanwhile, companies continue to post record profits while downsizing - gee, I wonder why.
The reality is, we need to start to figure out and move toward a new way to run the world. I always semi-land in a kind of Star Trek: TNG 'creation socialism', or whatever it is they have as a means of structuring a society, where you have a replicator that can create almost anything for you. They also have 'intellect machines' that are often used to build things from just describing what you want, or in other situation to dig into engineering problems, etc. Things that we are now starting to do with our new GPT tech.
Putting all this another way – how are all these corporations going to continue to exist when there are no customers to buy their products, because there are no jobs left. Basically, money as we know it is going to go away, we won't be a society of limited resources anymore.
And again as a reminder, this isn't something I'm saying is 2-5 or even 50 years away (maybe?). What I'm saying is that set the timeline however you want, it is going to happen, and we need to start to plan for it now. Based on what's happening already, we're likely at the start of a 20 – 50 year transition towards a new form of society entirely.
I doubt it. There will probably always be some scarcity. Even if we remove (or effectively remove) scarcity of physical resources or energy, there is still only a limited amount of time. You could base an entire economy on time and you'd still need some kind of a medium of exchange. Even in Star Trek TNG, every schmo on Earth doesn't have access to Federation supercomputing clusters or their own starship.
Most likely society will reorganize around it and new professions that don't exist today will be created.
And we haven't even considered the positive impact of AI.
It could for example accelerate the process of drug manufacturing and genetic therapies that will considerably increase human's lifespan, triggering another form of social reorganization.
So it is really impossible to tell with certainty how the world will look like in 50 or 100 years from now.
I am optimistic and would like to think it will be a much better place, not necessarily perfect and just in every aspect, inequality and armed conflicts will probably continue to exist, but overall it will be better than it is today.
There was a lot of fear during the industrial revolution, prompting many intellectuals to have a very grim outlook of the future, particularly when it comes to social issues (yes you, Marx).
But ultimately, if you look at the data compiled decades later, such as GDP and life expectancy around the world, it is undeniable that social and economic changes resulting from the industrial revolution made the world a much better place.
Not a perfect one, but certainly much better than it was (at least for humans).
When I show ChatGPT to a lay person they don't really care.
When I show it to a professional copywriter they say that if they submitted this content to a client they would lose the client.
I'm reminded of when my son was learning to talk and everything he said seemed brilliant and coherent to me.
To any stranger it sounded like gibberish.
I think GPT is like my son, and all tech people are like excited parents.
Maybe the kid will learn to speak like an adult, but it can't yet
But go back on HN a decade and be really honest.
Are we any better than average at forecasting?
I'm starting to think that being on the bleeding edge and seeing so many possible alternate futures actually makes us worse
Everyone is terrible at that.
I guarantee students are more excited about it than tech people because they are using it to successfully pass assignments.
I think the problem is that money saved from making those jobs redundant is not going back to the society but few heads on top. That's the fundamental problem, company moves money from payroll to tech that replaces the people and any savings are entirely company's profits.
And then they go ahead on tax avoiding spree so increased profits don't even flow back to the society
There are hundreds of thousands of people making their living on translating, and now most of them need to find something else to survive now.
The result of ChatGPT depends on the domain of the language. Technical reasoning based on many parts of input, is what the machine was designed for.. human language A to human language B depends on specific models built over the last ten years or so..
But if you subtract those two.. language structure and technical structure, you are left with giant holes in the way humans talk, interact and create.
Overall I would say that people who translate formal documents are in trouble yes.. people who translate to no-loyalty cost-sensitive markets like corporate ads for products or travel recreation things.. are in trouble. But there are many other kinds of communication, and therefore translation.
We are into a new reality and no need to lament the old one, because it is gone anyway
We keep committing seppuku by emerging technology to where hubris becomes a large scale financial or even a physical crash, and the smartest ones of us avoid the knee jerk reaction to quick over-adoption of new and unpredictable tech. The main problem is that there are so many forces pushing us towards adoption of whatever is packaged and marketed most heavily, and we need to incorporate better rollbacks and ways of gradually adopting new tech. Testing is also a quickly dying art, as seen with Twitter... This will be our downfall if not properly reigned in.
The people (self proclaimed tech leaders) who we exalt have already shown us that they are impulsive, and driven by greed, vanity, and ego. If we continue to let them act as "golden children", and continue to make them more and more wealthy, there will be no going back to fairness in our world. Tax them fairly, hold them accountable, stop letting them influence politics, stop electing rich people and people with conflicts of interest, and let everyone have a voice and equal opportunity to climb into responsibility based on their incremental successes... We're not doing any of that at all right now -- it greatly worries me.
This was true for Industrial Automation when machines automated physical work and humans could do knowledge work.
Now machines are on the verge of automating knowledge there would be no work left for humans.
Bad examples. Those are instances where you need human beings to provide interpretation of the context surrounding the translation/transcription, and where strict regulatory regimes are in place. Those are likely the last to be automated.
It's been close to two decades and I still wonder if that "pure" approach has any chance of ever turning into something useful. Except now it's not just language but "AI" in general: ChatGPT is not an AGI, it's a model fed with prose that can generate coherent responses for a given input. It doesn't always work out right and it "hallucinates" (i.e. bullshits) more than we'd like but it feels like this is a more economically viable shot at most use cases for AGI than doing it "right" and attempting to create an actual AGI.
We didn't need to teach computers how language works in order to get them to provide adequate translations. Maybe we also don't need to teach them how the world works in order to get them to provide answers about it. But it will always be a 80% solution because it's an evolutionary dead end: it can't know things, we have only figured out how to trick it into pretending that it does.
I guess the same intuition led to these new AI technologies...
And apparently (or so I heard, I think) feeding transformer models training data of Language A could improve its ability to understand Language B. So maybe there's something truly universal in some sense.
We're a bit more specialised than these new models. But that's it, really.
A lot of things just fall in place if you accept an obvious fact that the brain is just a powerful enough function with certain inputs and outputs.
Neurons are not that complicated.
There are interesting consequences here though: no stable grammars, no truly uniform thinking, no singular way to understand people.
So in this particular case of LLMs we've managed to optimise our way around having specialised brain structures using a powerful enough math function. Next steps require improvements in how the model is trained, how we can reduce the amount of training data, what additional machinery might be necessary, etc...
But, damn it, this very thing that makes humanity possible - our language - is solved now. Natural language is a solved problem now. That very thing that makes complex societies possible - it's done. This is the fact. And it is crazy.
EDIT: typos
(I did feel excitement while following the development of AlphaZero and its played Go matches, but that was because it was revealing greater depths and beauty in the human created game of Go. And I maintain some interest in following the development of self-driving, particularly by Tesla.)
With regard to LLMs I can see how they could be useful. I think more particularly useful when they work from a constrained corpus, so the user can know what they're drawing from (and thus the limitations of that knowledge base). The example site that been posted by its maker to HN [1] where you can ask questions against a particular book is a good one for showing the use of the tool I think. But it's just a tool and it's not in any way a breakthrough in our understanding of ourselves, of cognition or anything like that. I think the people who are making these claims can't distinguish science fiction from actual reality. They are fantasists and I think they are leading themselves and others into delusion.
Right at the beginning of the current wave (2010-2012) of ML approaches I did some work on ML systems and NLP, and back then I clearly saw how nothing truly outstanding is happening, we were only starting to figure our what GPUs were capable of.
So all of this was fun: NLP, ML, vintage AI. But nothing felt like it did was groundbreaking, or would solve fundamental true GAI problems, or was even close.
Yet, 10 years later, here we are. Language is solved. In most areas I know /something/ about (programming, ML, NLP, compilers) this is huge and makes mountains of knowledge obsolete.
I'm not saying AlphaZero was creative either. But because it was operating inside a system that was already beautiful and which had such a vast 'solution space' as you put it, its exploration into greater depths of that space I found intriguing.
I think that's the contrast for me. Machine learning can be useful and even intriguing inside constrained spaces That's why I liked AlphaZero, working inside a very constrained (but deep) space. And why I also find Tesla's progress with self-driving interesting. It's a constrained task, even though it has a huge range of variables. And again why I find ChatGPT potentially useful in drawing from a constrained corpus but still don't find the language it generates appealing. It comes across as exactly what it is - machine generated text.
It's how it interprets what people write and provides coherent answers. This was not possible previously.
AlphaZero, chess algos do not have to break this barrier, they work form a very clear and well-defined input. It was clear that a mixture of machine brute force and smart thinking would eventually beat us at these games. No magic here. Alpha family algos are /very/ understandable.
Language, on the contrary, is fundamentally not very well defined. Is it flawed, fluid, diverse... not possible to formalize and make it properly machine-readable. All the smaller bits (words, syntax, etc) are easy. But how these things come together - this can be only vaguely described through rigid formal frammars, but never fully.
Compare that to how on the lowest level we understand our brain very well. Every neuron is a trivial building brick. It's how super-complex functions of input to output arise from these trivial pieces - that's amazing. Every neural network is unique. Abstractions, layers of knowledge - everything is there. And it's kind of unique for every human so unknowable in the general case...
But it remains a tool and a derivative one. You will see people in these recent HN threads making grandiose claims about LLMs 'reasoning' and 'innovating' and 'trying new things' (I replied negatively under a comment just like this in this thread). LLMs can't and will never be able to do these things because, as I've already said, they are completely derivative. They may, by collating information and presenting it to the user, provoke new insights in the human user's mind. But they won't be forming any new insights themselves, because they are machines and machines are not alive, they are not intelligent, and they cannot think or reason (even if a machine model can 'learn').
I totally see your point about inherent "derivativeness" of LLMs. This is true.
But note how "being alive" or "being intelligent" or "be able to think" are hard to define. I'd work for the "duck test" approach: if it is not possible to distinguish a simulation from the original then it doesn't make sense to draw a line.
Anyways, yes, LLMs are boring. I am just not sure we people are not boring as well.
I agree, they are completely derivative. And so are you and I. We have copied everything we know, either from other humans or from whatever we have learned from our simple senses.
I'm not asking you to bet that LLMs will do any of those things really, I suppose it's not a guarantee that anything will improve to a certain point. But I am cautioning not to bet heavily against it because, after witnessing what this generation of LLM is capable of, I no longer believe there's anything fundamentally different about human brains, so, to me, it's like asking if an x86-64 PC will ever be able to emulate a PS5. Maybe not today, but I don't see any reason why a typical PC in 10 or 15 years would have trouble.
Just more pearl clutching and reverence for the mystical is what it looks like to me.
This is exactly the case. Religion is a constant in human society across time and space. And one of its main functions is exactly that - to keep people away from moral depravity. Speaking against this function of religion (again, proved across all cultures and all times) only shows profound shallowness.
Personally I'm not much of a fan of Christianity or the other Abrahamic religions. And I think generally its time is coming to an end (with a long tapering off). But I think you can see quite clearly that the moving away from Christianity (probably inevitable) over the last century or so has led to moral decline and depravity in the West. If Christianity's time is coming to an end, we will neeed (and I believe we will generate) a new religion to replace it. And I don't think it will be in the Abrahamic tradition.
So what you say is that humanity is not capable of learning moral behaviour without gods? And this is why it needs deities?
Religions are one of the traditional ways of motivating these rules... Legal systems are supporting them on an enforcement level.
I like religions as a very deep cultural phenomena/ideology having all kinds of effects on the, ehm, society. I just don't think religions are strictly necessary for a functioning society.
Not trying to poke holes, just clarifying.
Because for the other 20 percent it's plainly -not- good enough. It can't even produce an acceptable business letter in a resource-rich target language, for example. It just gets you "a good chunk of the way there."
And there's no evidence that either (1) throwing exponentially more data at the problem with see matching gains in accuracy or (2) this additional data will even be available.
Did people stop doing this at some point? Maybe after the advent of massively addictive social media, people more often ended up screenshotting it and sharing it for likes instead of correcting it.
I'm sure Google have stats on this but no idea whether they're public.
The fact we got this far through brute force is just insanely telling. This is a natural phenomena we're stumbling upon, not something crafted by humans.
Also - fun fact, the Facebook Llama model that fits on a Raspberry Pi and is almost as good as GPT3? Also basically brute force. They just trained it a lot longer and it shrunk the model. Food for thought.
However, translation of more distant languages is pretty terrible. Vietnamese to English is something I use Google translate for everyday and it's a mess. I can usually guess what the intended meaning was but if you're translating a paragraph or more it won't even be able to translate the same important subject words consistently throughout. Throw in any kind of slang or abbreviations (which Vietnamese people use a lot when messaging each other) and it's completely lost.
Really, words, utterances by themselves, carry meaning. Language is just a structure for _us_, so to speak, that we agree on for ease of communication. I think this is why probabilistic models do so well: the ideas we all have are mostly similar, it really is about just mapping from one kind of word to another, or kind of phrase to another.
Feel free to respond, I’m most certainly out of my depth here.
I've done a lot of work in NLP and the times when computational linguistics has been useful is very rare. The only time I shipped something to production that used it was a classifier for documents that needed to evaluate them on a sentence by sentence basis for possible compliance issues. Computational linguistics was useful then because I could rewrite mulit-clause sentences into simpler single clause sentences which the classifier could get better accuracy on.
> And here was Google Translate being "good enough" for 80% of all use cases using a "dumb" statistic model that didn't even have a coherent concept of what a language is.
I assume you are aware if Frederick Jelinek quote "Every time I fire a linguist, the performance of the speech recognizer goes up"?[1]
That was in 1998. It's been pretty clear for a long time that computational linguistics can provide some tools to help us understand language but it is insufficiently reliable to use for unconstrained tasks.
At the margin, these are equivalent (Chinese room, and all). I wonder if humans also learn similarly then retroactively tell themselves they actually do know things instead of just containing experiences encoded in their neurons (and whether that is any different than a neural network encoding trained "knowledge" in its neurons, too). This is the semantics of epistemology, at the end of the day.
Which of course is a good thing to make sure many people get to keep their jobs.
They fully understand that LLMs are stealing lunch money from established information retrieval industry players selling overpriced search algorithms. For a long time, my company was deluded about being protected by insurmountable moats. I'm watching our PR folks going through the five stages of grief very loudly and very publicly on social media (particularly noticeable on Linkedin).
Here's a new trend happening these days. Upon releasing new non-fiction books to the general public, authors are simultaneously offering an LLM-based chatbot box where you can ask the book any question.
There is no good reason this should not work everywhere else, in exactly the same way. Take for example a large retailer who has a large internal knowledge base. Train an LLM on that corpus, ask the knowledge base any question. And retail is a key target market of my company.
Needless to say I'm looking for employment elsewhere.
Can you link to an example?
Using a prompt like "Tell me how to build a graph database from scratch. Specifically, how to design the data model, implement the data storage layer, and design the query language." only gives a very vague answer. Sometimes it suggests using existing technologies.
Anyone know what I'm missing?
One of my initial prompts mentioned graph databases as an example of a scalable system, so I wanted to ask it about the design properties that make it so. I figured that because it was a book about designing systems, it could give me an outline of how a graph database works in practice.
It's pretty annoying how the site erases your prompt once you receive your output. By the time it finishes loading I've half forgotten what my original question was.
Since LLM’s can’t scope themselves to be strictly true or accurate, there are indeed good reasons, like liability for false claims and added traditional support burden from incorrect guidance.
Everybody is getting so far ahead of the horse with this stuff, but we’re just not there yet and don’t know for sure how far we’re going to get.
This isn't true though the techniques to do so are 1. Not as yet widespread 2. Decrease the generality of the model and its perceived effectiveness.
But in this example the AI could hallucinate a statement attributed to you it actually formed by putting together reddit comments.
Bing tries to solve this and succeeds somewhat. It will insert Wikipedia style citations against each of its claims. You can visit them and verify the statement if you want. And I do it often.
No reason why a future DocAI can't link to specific sections in internal documents whenever it answers a question.
Personally - I'm moving to more of a focus on analytical modeling. There is really nothing interesting about deep learning to me anymore. The reality is that any new useful DL models will be coming out of mega-teams in a few companies, where improving output through detailed understanding of modeling is less cost effective than simply increasing data quality and scale. Its all very boring to me.
One could look at the move from linear models to non-linear models or the use of ConvNets (yes I know ViTs exist, to my knowledge the base layers are still convolution layers) as 'leveraging human knowledge'. Only after those shifts were made did the leveraging of computation help. It would seem to me that the naive reading of that quote only rings true between breakthroughs.
We will always want to do the discovery ourselves, and I can see why fighting that instinct is a challenge for those in the field.
Clearly someone felt that there'd be a better inductive bias and attempted something else, and now CNNs are what's used "in the long run".
Most of those projects were the usual "solution looking for a problem to solve". Even those projects that might have had _some_ utility, would have been way more effective to buy/license a product than to develop an in-house solution. Because really, what's the use of throwing a dozen 25-30 years old with non-specialized knowledge, when there are companies full of guys with PhDs in NLP that devote all their resources to NLP? Yeah, you can pipe together some python, but these kind of products will always be subpar and more expensive long-term than just buying a proper solution from a specialized company.
To me it was pretty clear that those projects were just PR so that c-levels could sell how they were preparing their company for a digital world. Can't say I'm sorry for all the people working on those non-issues though. From the attitude of recruiters and employees, you'd think they were about to find a cure for cancer. Honestly, I can't wait for GPT and other productivity tools to wrech havock upon the tech labour market. Some people in tech really need to be taken down a notch or two.
That's an odd reason to want this.
While adtech, crypto and other bullshit gets massive funding because it can turn a profit.
The incentives to have a good society don't align with the incentives of financial capitalism.
I’m starting to see the term “tech bros” appear more and more in HN - before hand I more frequently saw it outside of this site.
Some people on HN I have seen really come down on those that use the term. I don’t.
Perhaps those of us in the industry ought to recognize that the term exists because of a growing resentment among people outside of the tech industry.
Your comment hints too as to why that is.
People start generalizing about groups like this when they've stopped caring about negative policy consequences which affect those groups. Politicians who blame wage stagnation on immigrants do not expect to have those immigrants who gain citizenship vote for them. Why do you think people might have stopped caring what happens to the group designated "tech bros"?
Still better to train lots of electricians and refrigerant techs and solar installers and all the other workers that the energy transition will need, of course.
Also have yet to Graeber's book ("Bullshit Jobs: A Theory").
Except for perhaps doctors (and even then residency is BS) all of those jobs are treated or paid like crap.
Doctors and nurses now spend more time entering data than talking to patients.
Teachers now spend more time entering grades into online systems and fielding messages from parents.
Not sure how tech is helping or hurting plumbers except for the standard GPS tracking that bosses use to follow them around.
Doctors and plumbers might make society work, but technology drives society forward.
My job is to optimize workflows which reduce cost, CO2 and make the world better and easier.
It's not my fault that society is unable to optimize everything for less work for everyone while everyone can have a good live.
those projects were just PR so that c-levels could sell how they were preparing their company for a digital world
This is exactly it. The 2017-2019 corporate version of "invest in AI" meant to build an in-house team to do ML experiments on internal data, and then usually evolved a bit to get some "ml-ops" thrown in so they could "deploy" the models they built. I spent some time with a few companies doing this and it always reminded my of "the cat in the hat comes back" when the cat let all the little cats out of his hat and they went to work on the snow spots... just doing busy work...Anyway it's a symptom of the hype cycle - AI was the next electricity, but there were no actual products and nothing clear to do with it, just hire a bunch of kids to act like they were in a kaggle competition, or worse a bunch of PhDs to be under-utilized building scikit-learn models.
Now that there are (potentially) products coming along that at least bypass the low-level layer of ML, having an internal team makes no sense. Maybe the most logical thing that will happen is the pendulum will swing too far, and this bubble will consist more of businessy types using chatGPT without remotely understanding it or realizing it's just a computer program.
Disagree. I was on one of these R&D/prototyping teams running ML experiments and you're right, it was the company wanting to present itself as future-leaning, ready to adapt, and I would say that at this point it was a good move to have employees who understand where the tech is going.
Companies with internal teams that are able to implement open source models are in a much better negotiating position for the B2B contracts they're looking at for integrating GPT into their workflow, they won't need GPT as much, if they can fallback on their own models, and they will be better able to sit down with the sales engineers and call bullshit when they're being sold snake oil.
You nailed it, although very few models actually ever got deployed to Prod at Fortune 500 non-tech companies and the few that did delivered little value. I'm a consultant and most internal AI/ML/DS teams that I interacted with were just running experiments on internal data as you said, and the results would get pasted into Powerpoint, a narrative created, and then presented to executives, who did little or nothing with the "insights". Reminded me of the "Big Data" boom a few years earlier where every company created a Big Data Team who then promptly stood up a Hadoop cluster on prem, ingested every log file they could find, and then..................did nothing with it.
I wonder if this is a bad as everyone thinks. When a new technology arrives which is not completely understood, isn't the right approach to try to find some applications for it? Sure, most will fail, but some valid use cases will likely emerge.
I'm pretty sure almost all technologies at some point were solutions looking for a problem to solve. Examples include the internet, the computer and math.
I think it is. If they actually do end up finding a problem to solve, that would be serendipitous but I imagine the vast majority of the time they find themselves in the business of trying to convince the rest of us to buy a thing that we don’t need. And while the latter may drive the economy to some degree as I get older I detest it more and more.
There have been no real advancements since the desktop model of the late 1990s. We might have more animations and applications running in virtual machines for security purposes, but literally nothing new has come out.
Even all the web apps are reimplementation of basic desktop capabilities from the decades before, but slower and with more RAM usage. They might be easier to write (I personally don't think so - RAD apps from the 90s were quicker to write and use) but the actual utility hasn't changed; if anything it's just shoving all of your data from your microcomputer to someone else's microcomputer, and being tracked and losing control of said data whilst you're at it!
And we have easier access to videos on the Internet, I guess??
It all seems to be missing the point of actually having a computational device locally. There is no computation going on. It's all digital paper pushing.
The problem with “stuff we don’t need” arguments is they are fundamentally nihilistic.
Everyone needs a flying car so let’s get on with it.
Also the Internet came out of DARPA which was a method of sharing data between geographically remote military facilities. It wasn't like they wired up devices and thought "what could we use this for?".
Only after DARPAnet solved that problem did it get adapted to some other problems (ex: how do I send cat pictures to people)?
I think the opposite -- nearly all technologies came about as a result of people trying to solve existing real problems. Examples include the internet, the computer and math. (Although I don't think "math" counts as a technology.)
The internet came about from darpanet, which was solving the problem of network resiliency. Computers automated what used to be a human job ("computer") of doing very large amounts of computations. That automation was solving the problem of needing to do more computations than could be done with armies of people.
You have to remember that when these sorts of things happen, the ones who get "taken down" in ways that actually affect their lives are invariably the ones who already have the least. The ones who "need" that takedown will be just fine, unless they've made incredibly stupid investment decisions.
I'm not sure that was the case with personal computing in 1980-s. What was the significant part of society which had the least and got "taken down"?
ChatGPT and other ML apps can find you the needle in the data haystack. To look up stuff on the PC you still needed to know the location of your stuff, filesystem info and how to formulate queries. You no longer need to learn to "speak machine language" but finally the machines can now understand human language to do what you tell them to do.
Of course, ChatGPT & friends can also say dumb shit or just hallucinate stuff up so you still need a human in the loop to double-check everything.
So it would be good to see the parallels with 1980-s here, if this generalization holds.
Yeah having a whole big team create the internal baseline is not cost effective, but having at least one or two people work on something to actually know the vendor is worth their cost is important.
In fairness most PhD topics people work on these days, outside of the select few top research universities in the world, are obsolete before they begin. At least from what my friends in the field tell me.
Counter-anecdata of one: On the other hand, one of the research teams of which I've been a member after my PhD was basically inventing Linux containers (in competition with other teams). Industry caught up pretty quickly on that. Still, academia arrived first.
edit Rephrased to decrease pedantism.
Could you give us more detail? It sounds intriguing.
All these things have been available in academia for a long time now. Even languages such as Rust or Scala, that offer cutting edge (for the industry) type systems, are mostly based on academic research from the 90s.
For comparison, garbage-collectors were invented in the 60s and were still considered novelties in the industry in the early 2000s.
Perhaps looking at the proceedings of ICFP and POPL can help?
A touch of understatement.
Their value just went up tremendously, even if their PhD thesis got cancelled.
Easily millionaires waiting to happen.
---
edit: Can't respond to child comment due to rate limit, so editing instead.
> That is not how it works at all.
Speak for yourself. I'm hiring folks off 4chan, and they're kicking ass with pytorch and can digest and author papers just fine.
People stopped caring about software engineering and data science degrees in the late 2010's.
People will stop caring about AI/ML PhDs as soon as the challenge to hire talent hits - and it will hit this year.
Hired in industry. That's the opposite. I've had a friend who had to hide that they had a PhD to be hired...
For a plain SWE role a Ph.d might be a disadvantage here too, but for anything ML related it is mandatory from what I can see.
As someone who hired for this in general we'd use PhD (or maybe a Masters degree) as a filter by HR before I even saw them.
It's true that a PhD doesn't guarantee anything though. I once interviewed a candidate with 2 PhDs who couldn't explain the difference between regression and classification (which was sort of our "ok lets calm your nerves" question).
Almost all the time, they're shitty startups, where bankruptcy is a matter of time, run by overpromising-underdelivering grifter CTOs pursuing a get-rich-quick scheme using whatever is trendy right now -crypto, AI, whatever has the most density on the frontpage-.
That was a few years ago, though.
The demand for AI/ML will fast outstrip available talent. We'll be pulling students right out of undergrad if they can pass an interview.
I'm hiring folks off Reddit and 4chan that show an ability to futz with PyTorch and read papers.
Also, from your sibling comment:
> Maybe it is also a matter of location. I am in Germany.
Huge factor. US cares about getting work done and little else. Titles are honestly more trouble than they're worth and you sometimes see negative selection for them in software engineering. I suspect this will bleed over into AI/ML in ten years.
Work and getting it done is what matters. If someone has an aptitude for doing a task, it doesn't matter where it came from. If they can get along with your team, do the work, learn on the job and grow, bring them on.
Chris Olah was at OpenAI but is now one of the founders of Anthropic but doesn't have any degree (he joined Google Brain after dropping out of his undergrad degree).
I think they are pretty exceptional though - most people on the ML team do AFAIK. Perhaps on the infrastructure side the backgrounds are different.
We've tried many time to work with CSIRO (the NSF of Australia) and it's fallen flat. They love impressive resumes and nothing else. I'm having a chat with their "Director of ML" who's never heard of the words "word2vec" or "pytorch" before. (And I'm a UX designer!)
I think at most corporate firms you'll end up running into more resume stuffers than people who actually know how to use ML tools.
Reminds me of Computer Security academic and professional pedigree where these people aren't even programmers
Will this effect the job market (both academic and commercial) for these folks? It's very hard to say. Clearly lots of value will be generated by the new generation of models. There will be a lot of catchup and utilisation work where people will want to have models in house and with specific features that the hyperscale models don't have (for example constrained training sets). I'm wondering how many commercial illustrators have had their practices disrupted by Stable Diffusion? Will the same dynamics (what ever they are) apply for the use of LLM's?
Pretty hard disagree. Even if your NLP PhD topic is looking at hypotheses on underlying processes about how languages work (and LLMs can't give you this insight), 9 times out of 10 it's with an eye for some sort of "applicability" of this for the future. GPT-4 just cut off the applicability parts of this for huge swaths of NLP research.
And aren’t PhDs supposed have a theoretical underpinning?
To be honest, that's a sh*tty PhD to begin with, it doesn't bring anything good and really worthwhile to the world, quite the contrary.
We could not "just use ChatGPT" because the answers have to be correct, or students could take legal action against the university in situations where they'd give an answer in an exam that the chatbot gave them and that might be wrong. It's exciting for sure to compare e.g. ChaptGPT with other approaches, for a set of controlled benchmark questions such as:
-- snip --
Q: What is the A* algorithm?
A: The A* algorithm is a popular heuristic search algorithm used in pathfinding and graph traversal. It is an extension of Dijkstra's algorithm, but uses a heuristic function to guide the search towards the goal node, which can make it more efficient than Dijkstra's algorithm in certain situations.
The algorithm works by maintaining a priority queue of nodes to be explored, where the priority of each node is the sum of its actual cost from the starting node (g-score) and an estimate of the cost to reach the goal node (h-score), based on a heuristic function. The heuristic function must be admissible, meaning it never overestimates the actual cost to the goal node.
At each step, the algorithm selects the node with the lowest priority from the queue and expands it, generating its neighboring nodes and computing their priority scores. The algorithm then adds these nodes to the queue if they have not been explored before or if a better path to them has been found.
If the goal node is reached, the algorithm terminates and returns the path from the starting node to the goal node. If the queue becomes empty before the goal node is reached, then there is no path from the starting node to the goal node.
Overall, the A* algorithm is a powerful and widely used algorithm that has many applications in fields such as robotics, video games, and logistics.
-- snip --Right now, GPT 4 would earn a top 10% SAT score, implying that it can give very high quality answers on a range of scholarly topic. This output is basically free.
Nobody that's beholden to even mild economic pressures is going to pay for an expensively constructed 100% solution if they can have the 99% solution for free.
My own opinion is that people are going to have to become creators. And quickly. You can still create digital products, but you'll need to be a lot more quiet about what you're doing. And you'll need to have a facility for abstract thought to come up with ideas that no one else has yet.
the only reason they studied, went to university etc was to avoid doing manual labour. this has been happening for decades, a century. they ll be depressed
They'll be depressed? Tough shit, we're all depressed.
Software dev. If computer steal my computer job, I have plenty of other physical skills.
Specialisation is for insects.
I was just curios.
This is part of what the original UBI concept was about.
If this doesn't happen, yes, there will likely be violence until it is fixed.
The other view is that many technologies that were supposed to reduce work actually net added work, because now more sophisticated tasks could be done by the humans, so the net was similar to the highway paradox where more and wider highways breed more traffic by induced demand.
Where would this demand come from? IDK, but at least initially, these LLMs make such massive errors that keeping a lid on the now-hyper-industrial-scale bullshit[0] spewed by these machines will make many more full time jobs.
Seriously, just today I was amazed at how the GPT model tried to not only BS me with completely fabricated author names for an article that I had it summarize, but it repeatedly did so even after being successively prompted more and more specifically to where it could find the actual author (hint: right after the byline starting with the word "Author". It just keep apologizing and then doubling down on more fantastic lies, as if it were very motivated to hide the truth (I know it's not, that's just how fantablous it was).
[0] Bullshit being defined as speech or writing telling a good tale but with zero regard to the truth or falsehood of any part of it — with no malice but nonetheless a salad of truth and lies.
Current "work for a living" systems only sustained the population because a human could be the most cost-effective way to get something done. Unless there are still tasks where human labor is the best option (research jobs maybe), this entire economic system will collapse.
But where will they make their billions from if everyone will be living on a basic income? Less money for them will mean less tax money, and less UBI. It will spiral out of control into complete societal collapse if AI doesn't hit a plateau soon.
That is why I said "solid UBI", as in more than merely survival wages, i.e., enough to not merely buy food & shelter, but also to live.
That said, this does need some thinking through multiple stages. On one hand, societies did still work when there was effectively unlimited slave labor, but that may be no more than a rough proxy.
Go to the endpoint assumption that AI and robots can produce everything needed for the population to live, and they are owned by 1% of the population. They made so much money so fast that they bought all the means of production. As of 01-January-2025 everyone is fired. Now what? Your'e right, no one can buy anything. The remaining populace cannot do anything because the new oligarchs have enough money and power to buy &/or threaten any politician.
The population overall is not going to simply lie down and die. About four days after running our of all the food in their pantries, they'll be revolting in the streets. One plausible result is a lot of carnage and the oligarchs are all killed and the AI is destroyed and outlawed. Or, they actually have sufficient command of the military and no defections and the military isn't smart enough to figure out that they're next on the starvation list, so the populace is wiped out as they revolt, and the world is left with the 1% of owners, and 1% of military. Or, there's some kind of balance reached, and the non-AI-owning class of fired people reconstitutes something similar to last year's economy, while the 1% go off to Mars or withdraw into their metaverse-ish thing.
That's just a few random thoughts on rolling the dice among the big forces, but it never plays out like that, so all of these are 99%+ likely to be wrong.
Seems like the only thing we know is that this potentially massively magnifies instability.
As mentioned elsewhere this is not the first technological disruption in the economy. The change from heavy industry to a service industry didn't go well, hopefully it's possible to take learnings from this and do it right this time.
For other legacy industries like the German ICE car industry it was at times a close call, so that's when the 35 hour week became widely adopted during the 90s. Even today it is still an option for new people joining (working on non-legacy products of course).
My ranking:
1. ChatGPT4 - flawless translation. I was blown away
2. DeepL - very close, but one mistake
3. Google Translate - good translation, some mistakes
4. Microsoft Translate - bad translation, many mistakes
I can understand the panic.
I guess we have to get used to software redefining the meaning of words. It was kind of funny when that happened regarding Google Maps / neighborhood names, but with LLMs it's a different ballgame.
For anyone who doesn't speak German, pathetisch means with pathos, impassioned.
A native English speaker probably would only use "pathetic" to mean "emotional" if the emotions were specifically negative. They also would use pathetic to describe someone experiencing non-emotional suffering such as injury or poverty.
Therefore, a native English speaker probably would not use "pathetic" to mean "emotional" in everyday writing. However, I could definitely see someone using it to mean emotional when they were being more poetic. For example, I could see someone calling an essay on the emotional toll of counseling "The Pathetic Class" in order to imply that social workers are a class that society has tasked with confronting negative emotions.
It's the same in Romanian, and I guess many other languages.
Many of the common words of European languages are derived from Greek and Latin, and where the meaning has diverged in English, now (because of its ubiquity) these false friends are being realigned to mean what they do in English.
And as with anything else, with the time it will get improved, too. LLM is not the answer to all linguistic problems.
https://github.com/ogkalu2/Human-parity-on-machine-translati...
But the interesting thing IMHO is the nature of the mistakes ChatGPT makes... often they're quite elementary mistakes (e.g., the occasional subject-adjective word order) while it gets the big picture right. Whereas DeepL is sometimes the reverse. ChatGPT also has the advantage of being able to tailor its output to a particular context, e.g., it can tailor legal translations to terms used in Canadian law rather than French law. However, for longer texts, I've noticed that ChatGPT will sometimes omit small parts of the source text from the translation, which is unfortunate.
I have a colleague who says that the Mandarin translations done by ChatGPT are an order of magnitude better than DeepL though, which is interesting.
All this training does not happen by itself.
But it's also extremely exciting, we'll be able to build really great things very easily, and focus our efforts elsewhere. Today anyone can throw together a language learning tutor to rival Duolingo. As long as you're in it for solving problems you shouldn't be too threatened by whatever tool set you're currently becoming obsolete.
OpenAI could build a state-of-the-art tool with a few hundred developers - to me, that means that money will converge to them and other big orgs rather than the opposite.
With a PhD in the domain, I consider myself pretty good at (a subset of) distributed programming. But these days, when companies hire for distributed programming, they seem to want developers who know a specific set of tools and APIs. I'm more suited at reimplementing them for scratch.
I see the AI stuff as very different from, say, the microcomputer revolution. People had LOTS of things they wanted to use computers for, but the computers were simply too expensive.
As soon as microprocessors arrived, people had LOTS of things they were already waiting to apply them to. Factory automation was screaming for computers. Payroll management was screaming for computers.
I don't see that with the current AI stuff. What thing was waiting for NLP/OpenAI to get good enough?
Yes, things like computer games opened up whole new vistas, and maybe AI will do that, but that's a 20 year later thing. What stuff was screaming for AI right now? Maybe transcription?
When I see the search bar on any of my favorite forums suddenly become useful, I'll believe that OpenAI stuff actually works.
Finally, the real problem is that OpenAI needs to cough up what I want but then it needs to cough up the original references to what I want. I normally don't make other humans do that. If I'm asking someone for advice, I've already ascertained that I can trust them and I'm probably going to accept their answers. If it's random conversation and interesting or unusual, I'll mark it, but I'm not going to incorporate it until I verify.
Although, given the current political environment, pehaps I should ask other humans to give me more references.
Smaller models trained supervised/in-domain are simply more efficient and more accurate than unsupervised/out-of-domain. Plus we own and operate the technology much more cheaply.
I don’t doubt that if your were trying to build a competing product to what OpenAI is doing that you’d feel affected, but there’s also a lot of other problems that are not being solved by generative models.
In the future -- forget about cosy job you can be doing for the rest of your life. You no longer have any guarantees even if you own the business and even if you are farmer.
What you absolutely don't want is spend X years at uni learning something, and then 5-10 years into your "career" finding out it was obsoleted overnight and you now don't have plan B.
That seems to be running directly opposite of the current trend of admin assistant jobs requiring 2 years specialized admin assistant diplomas. Tech (and I would guess the world of the business MBA) is a unique space where people are learning and changing so quickly, but for a lot of those outside the bubble things seem to be calcifying and requiring more and more training at the expensive of the worker.
Extremely relevant story
I think everyone mostly agrees that AI is coming for a lot of jobs. There's disagreement about how many, how it will impact society and the like.
The pace of technology is not linear, it accelerates. I've never seen something that has so rapidly crossed into the "magical" territory as "nearly every single big LLM/Generative AI thing" seems to. It redefines what was previously laughably impossible ... a decade ago.
We're riding a curve upward that is making it extremely hard to see what's coming next. All of the pontificating, all of the attempts at finding solutions to imagined problems ... I can't see one that doesn't feel like a blindfolded person aiming at what they were told was a dart board with what they were told was a dart. There's really nothing to do but hang on and hope you land where any new opportunities creep up.
Expect bubbles, black swans, and purple unicorns.
We have had some issues and complaints with the API,( mostly with GPT 3 as the fine tuning was only open for the base model and that had some trouble with some questions). Also there is a finicky response time, despite using having paid access. Response time varies from 10 seconds to even a minute (during some downtime that occured a few days ago, and a few days even before that there was a complete outage).
Everyone is going to feel this, most prominently people in the sorts of industries that frequent HN. If you haven't yet, you will or you will be forced to when you discover everyone in your field is out-producing you armed with these tools.
How's them NFTs and Blockchain doing the watershed world changin these days?
(2009) https://www.computerworld.com/article/2525229/study--interne...
(2015) https://venturebeat.com/mobile/mobile-technology-has-created...
(2021) https://www.iab.com/news/study-finds-internet-economy-grew-s...
Same for the internet - things changed, but the breathless predictions that retails stores are dead and everyone will be WFH haven't happened. It became a tool.
So will ChatGPT - I look forward to good autocomplete and a nice bullshitter-in-a-box tool to generate these HN comments. But most jobs aren't improved or more productive by adding a faster bullshit generator. The bullshit generators in suits are already really good at that.
db112nl@gmail.com Why wasn't Blockchain a watershed moment but chatGPT is.
There are several factors that contribute to why blockchain and chatGPT have had different impacts on their respective fields.
Firstly, blockchain technology has had a significant impact in certain areas, particularly in the financial industry, where it has been used to create new forms of digital currency and facilitate secure and efficient transactions. However, it has not yet achieved the widespread adoption and transformative impact that some had predicted. One reason for this may be that the technology is still relatively new and complex, and many organizations are still trying to figure out how best to implement it.
In contrast, chatGPT is a language model that has been designed to understand and generate natural language text. This technology has been more widely adopted because it has a broad range of potential applications, from customer service chatbots to language translation services. Additionally, language is a fundamental aspect of human communication, which makes the potential impact of language models like chatGPT more readily apparent to people.
Another factor that may have contributed to the different impacts of blockchain and chatGPT is the level of public attention and interest. While blockchain has received significant media coverage, it has not captured the popular imagination in the same way that chatGPT has. ChatGPT has been widely publicized and has even been featured in news articles and talk shows, which has helped to raise awareness and generate interest in the technology.
In conclusion, the differences in impact between blockchain and chatGPT can be attributed to a range of factors, including the complexity and novelty of the technology, the level of public awareness, and the broad range of potential applications.
I'm sure that for the people in the thick of crypto and web3, it would've felt like that was the most important thing on everyone's mind.
Why the hell stay in in academia? This is clearly the next technological wave, and you shouldn't sleep on it. Especially when you're so well positioned to take advantage of your experience. You can make $500,000/yr (maybe more with all the new startups and options) and be on the bleeding edge.
If you want to go back to academia later, you can comfortably do so. Most don't, but that doesn't mean it isn't an option.
ETA: And though it may take longer, people who understand these models will eventually be in possession of the most valuable skill there is. Perhaps one of the last valuable human skills, if things go a certain direction.
Getting your hands dirty is the best way to understand how something works. Think about all the useless SE and PL work that gets done by folks who never programmed for a living, and how often faculty members in those fields with 10 yoe in industry spend their first few years back in academia just slamming ball after ball way out of the park.
More importantly, $500K gross is $300K net. Times 5 is $1.5, or time 10 is $3M. That's pretty good "fuck you" money. On top which some industry street cred allows new faculty to opt out of a lot of the ridiculous BS that happens in academia. Seen this time and again.
I think the easiest and best path for a fresh NLP phd grad can do right now is find the highest paying industry position, stick it out 5-10 years, then return as a profess of practice and tear it up pre-tenure (or just say f u to the tenure track because who needs tenure when you've got a flush brokerage account?)
$100,000 in 1970 is worth almost $800,000 today.
Yes, downvote me all you want. But if you're an NLP expert thinking of working for a company that will make billions off your work, you can and should demand millions at least.
Where is some evidence that NLP is 'solved'? What does it even mean? OpenAI itself acknowledges the fundamental limitations of ChatGPT and the method of training it, but apparently everybody is happily sweeping them under the rug:
"ChatGPT sometimes writes plausible-sounding but incorrect or nonsensical answers. Fixing this issue is challenging, as: (1) during RL training, there’s currently no source of truth; (2) training the model to be more cautious causes it to decline questions that it can answer correctly; and (3) supervised training misleads the model because the ideal answer depends on what the model knows, rather than what the human demonstrator knows." (from https://openai.com/blog/chatgpt )
Certainly ChatGPT/GPT-4 are impressive accomplishments, and it doesn't mean they won't be useful, but we were pretty sure in the past that we had "solved" AI or that we were just about to crack it, just give it a few years... except there's always a new rabbit hole to fall into waiting for you.
I've been asking it about lyrics from songs that I know of, but where I can't find the original artist listed. I was hoping chat gpt had consumed a stack of lyrics and I could just ask it, "What song has this chorus or one similar to X..." It didn't work. Instead it firmly stated the wrong answer. And when I gave it time ranges it just noped out of there.
I think If I could ask it a question and it could go, I've used these 20-100 sources directly to synthesize this information, it'd be very helpful.
https://dkb.blog/p/bing-ai-cant-be-trusted
To answer the question above, these systems cannot provide sources because they don’t work that way. Their source for everything is, basically, everything. They are trained on a huge corpus of text data and every output depends on that entire training.
They have no way to distinguish or differentiate which piece of the training data was the “actual” or “true” source of what they generated. It’s like the old questions “which drop caused the flood” or “which pebble caused the landslide”.
> Their source for everything is, basically, everything. They are trained on a huge corpus of text data and every output depends on that entire training.
Bing chat is explicitly taking in extra data. It's a distinctly different setup from chatgpt.
I think LLMs have essentially solved the natural language processing problem but they have not solved reasoning or logical abilities including mathematics.
ChatGPT cannot even reason reliably on what it knows and doesn’t know… it’s the library of Babel, but every book is written in excellent English.
Knowledge representation is a separate problem. NLP gives us some insights into what works here, but the multi-modal aspects of things like GPT4 show there is a lot more to knowledge presentation than just NLP.
LLMs produce perfectly fluent output and can understand natural language input as well as any human.
However knowledge representation is not solved. We still don't know how to interface a perfect LLM to other systems in the same way a human does things like looking up facts we aren't confident of or using a calculator to do math we cant' do in our head.
These are very significant problems and super important. But they are more adjacent to NLP in the same way tasks like something like Text-to-SQL [1] isn't a pure NLP task.
[1] for example https://github.com/salesforce/WikiSQL
Playing with Llama 65G gave me a sense for what the median raw effort is probably like. It seems to take a lot of work to fine tune and harness these systems and get them reliably producing useful output.
Bad analogy- if you had an integrated circuit team in your product company building custom CPUs and Intel came out with the 8080 (or whatever was the first modern commercial chip), probably time to disband the org and use the commercial tech
More recently Google, Microsoft, Apple, etc. decided they wanted to have speech recognition as an internal piece of their platforms.
Google poached lots of Nuance's talent. And then Microsoft bought what remained of the company.
Now speech recognition is a service integrated into the larger tech company's platforms, and also uses their more statistical/ML approaches, rather than being a component created by specialist companies/groups.
(I'm sure I'm grossly simplifying this — just seeing a potential parallel.)
I've never really had the gear or the skills to put together anything that improves over what I can pull from huggingface.
What I do have, and virtually none of my (not remotely technical) colleagues have, is a clue what to do with all this stuff.
They reckon it's about churning out poems and boilerplate text, the minute I figured it could give me whatever json I could reasonably ask for from a source doc, I was overjoyed.
I see more things I can be doing now, not a risk of being replaced.
For example, back in the 60's my dad was working on his book. The text was typed out double spaced, and he (and others) would make corrections. After a while, my mom would retype the whole thing.
Imagine typing a whole book. Again and again and again. She'd type hour after hour. It's dehumanizing.
And then came word processors. What a magical revolution! You could edit text instead of typing it all over again. I bet few people today realize what a great achievement that was.
All chatgpt does is select the most likely next word out of a corpus of existing text. It is not creative.
We don't need rooms full of typists anymore. Good riddance. I bet we get rid of a bunch of drudgery jobs with chatgpt.
I'm not sure but I'm now curious as to what the execs there are thinking, especially now with the recent Microsoft 365 news. Feels like the body blows keep coming.
Things that were a struggle 5 years ago are about to be easy.
so finally the tech sector is experiencing themselves what they have done to other lines of professions for the past decades, namely eradicting them (rightfully) with innovation?
well same advice applies then:
* embrace, move on and retrain for another profession * learn empathy from the panic and hurt