An Adversarial Review of “Adversarial Generation of Natural Language”
medium.com
medium.com
This newsletter looks interesting (mentioned on Twitter): http://digest.deeplearningweekly.com/issues/deep-learning-we...
This seems like the only real discussion from the search: https://twitter.com/kchonyc/status/873305485428088833
I guess search is broken, the author's twitter feed has a bunch more discussion (https://twitter.com/yoavgo/with_replies):
https://twitter.com/yoavgo/status/872831207163265024 https://twitter.com/yoavgo/status/872968874521702400 https://twitter.com/Smerity/status/872971766959718400 https://twitter.com/yoavgo/status/873489277157507072 https://twitter.com/jacobandreas/status/873109327644573696 https://twitter.com/yoavgo/status/873175136056336386 https://twitter.com/hugo_larochelle/status/87336968867543040... https://twitter.com/kchonyc/status/873306255204507648 https://twitter.com/yoavgo/status/873786844315607040 https://twitter.com/haldaume3/status/873565061754781697
My definition of "works" is https://en.wikipedia.org/wiki/DWIM, others are free to disagree.
Most researchers/academics lie somewhere on this spectrum. (Well I guess most human beings involved in any activity probably).
On the one end are salespeople who love to make a mountain out of a molehill they just discovered. On the other, slackers are like the perfectionists who never get anything done because they never resolved their analysis-paralysis.
There are very few who are exactly in the middle of the spectrum. The middle is a point of unstable equilibrium. You have to work very hard to stay there and can easily fall off to one side or the other.
This is especially true when two products have teams of about equal size, experience, and expertise. Time spent selling is time not spent making a better product. All other things being equal, every hour spent selling is one less hour spent working on making your product better.
This doesn't work in general, though.
> All other things being equal, every hour spent selling is one less hour spent working on making your product better.
Which, I believe, is a key to understanding why many (most?) of the things you can buy are utter crap, barely fit for the purpose they were made (if at all). It explains why so many successful SaaS businesses offer barely functional products. Because every hour spent selling is a hour spent not working on a product, and marketing has much better ROI than actually building something useful.
And of course you can sell a product that doesn't even exist.
The dimension is confidence or maybe approval.
You can be a DK without any pressure. With research/academia,
- there is a lot of pressure to excel and show that your work is the next best thing since sliced bread. Another popular phrase is 'publish or perish'
- the audience of such researchers are burdened with information overload, and that adds to the difficulty of having your voice heard, leading to a habit of taking shortcuts and trying to wow the audience when you really didn't do much.
Right now many people are engaged in creating blizzards of papers, the mechanism of choice is the construction of factories for writing papers - an army of graduate students and post-docs - and the construction of communities of publication that are plausible enough to enable the extraction of funds from funding sources.
Is this "useful" well - there is still bias and discrimination; there are not enough women (where enough is proportionate to the representation of women in the general population and their desire to do this kind of thing) and there are not enough people of colour - but there are more, which, thank god, I think most people think is a good thing. Research in some parts of CS is going very fast now as well, which is great!
But, in HCI, Enterprise IT, Software Methodologies, things do not seem so good, and where things are going well, like in AI I wonder if this is coincidence (as in look, a thousand cores have arrived, I can do many things while waiting for my paper factory to make more papers).
And it is expensive, very expensive. Where-as old CS involved a crowd of poorly paid sports jacket wearers, new CS involves a horde of hoodies, but the expense that worries me is the cognitive one - how to sort through the morass of chatter that conceals (intentionally, often, so as to enable the next seven or eight six page koans of review passing flim flam) the things I need (TM).
So, some things are better in CS land. People who wouldn't have got a shout before are in with a shout (some of them) but instead of building an inclusive community of people really trying to do research we have build a thing that I hesitate to call a community that is more inclusive but contains many people who are doing things which are not research at all, in fact, they are kinda anti-research.
And you can't get funding to do a field study, or publish a paper without waving some silly maths that means almost nothing about. And if you do publish, you'll have to cough up £2k one way or another, and no one will read the bloody thing.
As do most of the Google Translate pieces, even though I get the feeling that automatic translation of texts is now seen almost as a solved problem (it's not): all that Google Translate does is change some text from some original language to a second one, which is not a real language, it's just a language that's sometimes very close (grammatically and lexical) to one which the agent/user knows.
The idea is that we should try to look harder and have fairer judgements about the actual results and not get stuck on the methodologies.
Anyone else thought that this was very weird? The author appears to be complaining about the fact that reputable people/labs can post a PDF on arXiv and be taken seriously. How is this avoidable? Without arXiv, they could just post the PDF on their website or anywhere else.
The "risk" associated to publishing crap on arXiv is the same as always: have people notice it's crap and get a bad reputation. I'm not sure what ideology has to do with it.
The issue is "flag planting" : arXiv is one way to do it; the same arguments hold true for other sources but arXiv is used as a canonical example for "flag planting".
We have reached a point where reviewers of reputable conferences will ask why a paper is not referencing unreviewed work that has been uploaded on arxiv a day before the conference deadline.
This is not hyperbole, as anyone who is submitting to ICLR or NIPS can confirm. Work of certain labs is taken as established authority as soon as it hits arxiv.
On the other hand, arxiv has a huge audience, lately maybe even more than DL/NLP conferences, making the flag planting really effective, especially if you are from a prestigious group. So there is a real problem now with large, prestigious groups posting half-assed preliminary results in arxiv, and deterring more modest groups from working on the problem or bogging them down because they now need to compare themselves to the well-known arxiv approach, which often has serious reproducibility issues because it has not been peer-reviewed.
I'm not an arxiv hater, in fact I check arxiv every day and for the last year or so I have posted most of my papers in it. But the problem is real and something must be done. Not about arxiv which is just the messenger (and a good one), but about the flag-planting culture using it that has emerged in the field.
I am working on a social network for papers that addresses this problem. Arxiv is too vulnerable to fake news articles because it is an archaic social network. I also think https://openreview.net can be improved. You can check my project here : http://www.startcrowd.club it does not have an anti fake news feature yet, but it is on the product roadmap.
It appears that we agree that the problem is not with arXiv. Part of the problem is unsolvable: if prestigious groups can announce what they are working on and discourage other groups from working on the same problem, this may just be reasonable self-interest from the smaller groups and can hardly be avoided. As for reviewers asking for comparison to these works, I guess the problem lies with the reviewers: if an arXiv preprint has not been refereed and/or is hard to reproduce, it should be OK to say so in another paper and not be blamed for the lack of comparison.
In any case, this is an interesting problem, thanks again for making me aware of it.
"Let the market decide!"