HNHacker News
TopNewBestAskShowJobs

gradys

1,341 karma · joined March 25, 2012

grady dot hsimon on Gmail

@GradySimon on Twitter

submissionscomments
gradys··on A New Lens on Understanding Generalization in Deep Learning
But maybe with 5 million real examples sampled from the distribution CIFAR-10 was sampled from they would in fact see a difference. Maybe the generative model is capturing only a limited slice of the diversity that the ideal model would really see.

It seems like they should have downsampled from an actually large dataset rather than generatively upsampled from a small dataset. Unless I'm missing something?

gradys··on There Should Be No Computer Art (1971) [pdf]
The computer as a tool for an individual to use is a perspective that comes naturally to us now, but note that this article was written before even the Alto and the Apple II, and a decade before the IBM PC. Interactive computing was a fringe idea that only a relatively small set of people were thinking about. Computing meant batch computing.
gradys··on Amazon Sends ‘Vote No’ Instructions to Unionizing Employees
Do you think there's any chance that this was actually what a majority of Californians wanted? You seem very confident that you know better what the voters wanted than they did.

While differences in spending surely have some effect, spending isn't the only way to get the word out. I didn't measure, but I feel pretty confident that no more than 60% of the total communication about Prop 22 that I heard from any source was pro. I think it could have been a minority actually. For every PR firm the pro side hired, there were thousands of unpaid activists on the con side (which does say something positive about the con side!).

I noticed you brought out the "lawmakers can't repeal" line. This is true, but also effectively standard for ballot measures in CA. In fact the default is that you can't repeal or change the law under any circumstances. This was one point that initially swayed me to vote "against", but once I learned that this was basically how all ballot measures work (including ones you support, I'm sure), I swung back to voting "for".

For me, this was a difficult call, but I voted "for" on the basis of the net good for the entire pool of current drivers, admittedly at the expense of the smaller number of drivers who would have been able to drive under a better deal had this passed.

The net effect this would have had on driver welfare is far from obvious, so it always bugs me when people assert that the only reason this passed by a 17 point margin is that people were tricked and confused by the ride sharing companies.

gradys··on Assigning Blame for the Blackouts in Texas
My understanding of the term "neoliberal" doesn't include relentlessly lowering taxes or shrinking government. Maybe you mean "neoconservative" or "libertarian"?

To me, neoliberalism is better characterized as balancing market and non-market approaches to addressing social problems (more market-oriented than progressives, more comfort with government intervention than conservatives).

Also, sorta weird to blame Silicon Valley attitudes for a failure of the Texas government. Yeah a couple of companies from SV are opening offices in Texas, but this problem has been at least a decade in the making and it's not like SV is a dominant force in Texas.

gradys··on Google Open-Sources Trillion-Parameter AI Language Model Switch Transformer
Yep, it would definitely be difficult to justify running it in production. Accuracy would need to be higher as you said or it would need to be applicable to more tasks such that you can take other models out of production.

This kind of model could be used as the teacher in a distillation setup too though. Then faster training of the teacher is actually a huge benefit since it speeds up model development iteration cycles.

But even if it weren't practical to use in production in any sense, I'd argue there's value in doing the basic research of exploring design space of architectures in this way. This came out of a research team at Google. It may inspire and inform smaller, more practical architectures.

gradys··on Google Open-Sources Trillion-Parameter AI Language Model Switch Transformer
One key idea here is to use a very large number of parameters (model weights), but only use some subset of the parameters on each example. The parameters are divided up into blocks called "experts", and then some subset of experts are used on any given input. Which subset is used is chosen by the model itself in a data-dependent manner. This can be thought of as letting the model specialize different experts to handle different situations.

The advantage, as they show, is that the model can train to a given level of performance much faster with a fixed amount of computing power compared to an architecture that uses all parameters on every step. This might be because it allows you to have a very large number of parameters that can store a lot more specialized information without incurring as much of a computational cost. Of course the downside is that you end up with a very large model that literally won't fit in a lot of environments.

gradys··on Google Open-Sources Trillion-Parameter AI Language Model Switch Transformer
Mixture of experts architectures like this are a specific design decision to increase the parameter count but use and update those parameters sparsely. Sure, that design decision doesn't fit all scenarios, but it fits some, and it has its own advantages, like faster training time.
gradys··on Science fiction hasn’t prepared us to imagine machine learning
How do I make sure I hear about Invisible Sun when it comes out?
gradys··on Jeff Bezos Pitching Amazon.com (1997) [video]
I guess maybe I'm alone then, but I don't hate using Amazon.
gradys··on Factorio 1.1 stable
That optimization makes the 3D games more complex, not less. What Factorio does is impressive and surely comes with its own challenges, but it isn't necessarily more complex.
gradys··on How r/WallStreetBets gamed the stock of GameStop
Any evidence to back up the bot speculation? Normal social media memey exuberance seems sufficient to explain this to me.
gradys··on Controlling GPT-3 with Logit Bias
This is a great tip, thanks! Is this essentially importance sampling? Reminds me of a technique for off-policy RL.
gradys··on Show HN: Plant-based flavor enhancer – hack your tastebuds to eat healthy
Sorry, my question wasn't very clear. I guess what I meant to ask is whether this works when eating food that doesn't have as as much acid as a straight lemon. Sounds like the answer is yes.

Also, tangential, but I think the advice to brush your teeth right after is not right. Not a dentist, but IIUC, after exposing your teeth to acid, the enamel is in a weakend state and then the abrasion from brushing will do more damage. Rinsing with water would be helpful. Brushing before can also be helpful since the fluoride temporarily protects the enamel.

gradys··on Show HN: Plant-based flavor enhancer – hack your tastebuds to eat healthy
With the party trick version of this (eating lemons or limes straight), I was always concerned about the acidity on my teeth. Is that not a concern here because you just don't use as much acid?
gradys··on Google Images Restored
I guess the question here is who is more like those doctors in this scenario: you, or the people running the randomized controlled trials?
gradys··on Ask HN: What are you working on?
I made a tool for visualizing text datasets in 2D or 3D: https://nebulate.ai

It runs a machine learning model in your browser to convert the text into points in a high dimensional space, and then it projects those points down to 2/3D.

Right now you can tell it to visualize post titles or comments from any subreddit or load an hourly updating snapshot of Twitter.

You can also view your own data in it by selecting the New Nebula option. The data never leaves the browser, which also means the ML models are run in-browser (via tensorflow.js). This part might be slow and only works in Chrome unfortunately.

If you're interested in this kind of thing, I'd love to hear from you! Here or by email (grady.hsimon at gmail)

gradys··on 2020 Mathematical Art Exhibition
A Twitter account with this vibe: https://twitter.com/dmitricherniak

Would love to hear about others.

gradys··on Ask HN: Who's looking for a co-founder?
I'm a machine learning engineer with a focus on data efficient NLP. I want to build tools that make it dramatically easier to apply ML (especially NLP) to real world problems like:

- Routing customer support requests - Understanding freeform user feedback - Understanding the memetisphere of social media - Automating content moderation on social platforms - Bringing order to large document archives

I think it's possible to build a code-free, interactive interface that enables all of these things (though it may be best to focus on a single vertical).

Hit me up if you're interested in any of this. grady.hsimon at gmail.

gradys··on Elon Musk moves to Texas
If you mean 2008, yeah, you're right, but those were mostly loans to relatively low income people buying homes, not loans against billionaires' unrealized capital gains.
gradys··on Elon Musk moves to Texas
Why does taking private loans against his private assets count as socializing risk? The banks he gets these loans from understand the risk and do not represent "society".
gradys··on Ask HN: What's the best paper you've read in 2020?
Thank you! We have very similar interests! Especially the first one and the interactive sound recognition one. Any other work in IML you'd recommend?
gradys··on Join 'The Best Year Club' and commit to making 2021 the best year ever
I had a similar reaction, and also signed up. Upon reflection, I think there actually is a chance for it to be the best. Global poverty continues to shrink. Depending on your politics, you might be excited about the next US administration. Vaccines look like they'll be pretty widely deployed. Maybe we'll have a few more people out there doing good deeds.

Things are looking up!

gradys··on Show HN: Open-Source Memex – Alternative Approach to Roam/Obsidian
I'm super interested in this area, with my own vaporware attempt at building it (fully abandoned, unlike yours).

I'm an ML engineer focused on NLP applications. Contact info in my profile if you ever want to chat, e.g. about different approaches for estimating document similarity.

gradys··on A History of Clojure [pdf]
IMO, the ClojureScript SPA ecosystem is a stronger competitor in its domain than Clojure on the backend.

Re-frame in particular is a gem. It's as though someone tried the React/Redux stack, thought long and hard about actions and selectors, and realized that with one or two more pieces, everything falls into a beautiful, purely functional harmony.

gradys··on Demo of an OpenAI language model applied to code generation [video]
I worked on project very much like this last summer, a transformer language model applied to code completion.

You'd be surprised how easy it is to get a model that performs as well as what you see in the video. And it's even easier now that people have built great libraries for fine-tuning generative language models.

I encourage you to try it yourself! There are many interesting extensions for people to explore:

- Use bi-directional context (vanilla GPT-2 only sees backward context)

- Integrate with semantic analysis tools.

- Experiment with different context representations. You condition the model on an arbitrary sequence of N tokens. It's not necessarily the case that you should spend that whole budget on the N tokens that came immediately before. What about including the imports at the top of the file? What about the docstrings for functions that were just used? What about the filepath of the current file?

Don't look at something like this as though watching your job be automated away. Look at it as a tool that you can master and use to move up the stack.

gradys··on Demo of an OpenAI language model applied to code generation [video]
I'd put it differently. This is going to take your job, just like an assembly programmer from the 70s might consider Python to have basically taken their job. In software, the job is constantly eating itself and transforming.

It's part of the job to continually incorporate new capabilities and lever yourself up.

gradys··on Demo of an OpenAI language model applied to code generation [video]
Curious what analysis you're referencing here.

While I don't doubt people have shown that various transformer models have certain limitations, I'm pretty bullish on transformer models in general.

Here's a post exploring the application of transformers to symbolic mathematics for instance: https://medium.com/analytics-vidhya/solving-differential-equ...

gradys··on Ask HN: Dark mode for HN please?
I'm late to this party, but this is fantastic. I've popped it into the Firefox Stylus addon. Consider me a happy user. Thanks so much!
gradys··on A Brief History of California’s Housing Crisis
> Tech companies unnecessarily demand butts in chairs in the Bay Area, foisting the externalities onto the commons (infrastructure and housing).

I'm not so sure it's unnecessary. I (and my employer) find it very useful that I have so many colleagues within walking distance of my desk and that we have centralized meeting, event, and dining infrastructure.

I also find it very useful to have so many other employers nearby that I could work at if this job doesn't work out. My employer finds it very useful to have such a large local labor market for the kinds of roles they need to fill.

I'm not opposed to income and property taxes on these marginal increases in productivity that the tech industry gets in the Bay Area, but I strongly disagree that the industry's large presence is somehow inefficient. The tech industry actually uses the resources of the area more efficiently than other industries because of these network effects.

gradys··on Google to Samsung: Stop messing with Linux kernel code. It's hurting Android
The highest level of severity is reserved for anything that allows users to uninstall the crapware they ship with their phones.
← PreviousPage 4 of 10Next →