Zillow lost money because they weren't willing to lose money
stevenbuccini.com
stevenbuccini.com
I totally agree. It's not impossible to imagine their model working: why couldn't you serve as a market-maker for homes at a large scale, especially with the unique insights Zillow could have based on their datasets.
However I think where the hubris lay is in how they thought they could leapfrog all the way to an automated solution before building a competency as a house-flipping company.
From what I understand, where they failed was partly in building a rich enough model to properly account for the less easily quantifiable elements which ultimately account for a property's value. I.e. the price per square foot might make a property look like a steal, while something like a sewer main nearby, or problematic neighbor could radically change the value proposition to anyone standing at the site. That's a non-trivial problem to solve for even the best ML and it's not clear how you would automate this.
If you ask me, instead of focusing on building an automated price discovery system, they should have started by trying to build a quality home-flipping organization, and figuring out how to super-charge manual work using their datasets. Over time you might find ways to optimize the process and increase the level of automation to scale output relative to head-count.
If you pledge to purchase at the Zestimate then people who reasonably think they can get more than the Zestimate on the open market don’t have an incentive to sell their house to Zillow (besides convenience). But people who think the Zestimate is an over estimate will of course sell to Zillow. So instead of a normal distribution of actual value:estimated value you end up with a skew towards the end where the estimate is over the actual value.
Trading housing is very different from normal market making because houses are not fungible commodities like most securities are. For most entities trading securities at low frequency it does not really matter whether a market maker skims off a few pennies on their trade; it’s worth it for the liquidity. Houses are less liquid (because they are non fungible) so the liquidity is more valuable, but the price improvement routing around a MM can also be many percentage points of a trade because there are not only so many factors affecting their valuation, but also just chance and random noise (bidding war, a particular buyer falling in love with the property, not-price-conscious buyers).
According to Matt Levine's recent column, while you might think that, it wasn't what sunk them in practice. Bidding low in fact worked; it just was inherently limited in scale, which is why they switched to bidding higher. Unfortunately, being wrong in the other direction is very bad.
"I know, I know, the traders are saying: “No, this is stupid, your algorithms will not be 100% precise, some of your ‘lowball’ bids will in fact be too high, and those will be the ones that sellers accept. You’ll get adverse selection and end up losing money.” But that was not Zillow’s actual experience in the first quarter! The actual experience is presumably that some people accidentally got too-high bids, realized they were good and accepted them, but mostly Zillow sent too-low bids to everyone, and some people, for whatever irrational reason — market ignorance or financial necessity or laziness or whatever — accepted the too-low bids. The general point is that there is no reason at all to think that the people on the other side of these trades from Zillow are generally better informed than Zillow is. Sure they know more about their houses than Zillow does, but Zillow knows more about the market, and has more money"
"If you systematically bid too low, you will not do many trades, but you will make a lot of money on each trade. If you systematically bid too high, you will lose money on each trade, and also you will do a whole ton of trades. This is much worse!"
On the other hand, for a house that you could sell "instantly" for, say, $450,000, but that you could potentially get by selling privately for $480,000 or $500,000, that is now leaving 10 times more dollars on the table. Proportionally, the difference between the car and the house might be similar but in absolute dollars, it's a huge difference.
I mean, why did Zillow go straight from making profitable purchases of few houses to losing money on tons? If there's some sweet spot in the middle that would have allowed them to scale up profitably, would it really be that hard to land on that spot?
I think it's obvious that homeowners know much more about the houses and neighborhoods they live in than Zillow, and this creates a big risk for Zillow.
When they were paying less than market price, they still made a lot per deal, and when they were paying more than market price, adverse selection hardly pushed them over the edge.
If you make "a million high ball offers" then the fact that some of them are particularly bad deals isn't the core problem.
I'm not 100% confident of this but it seems like the picture that is being painted.
And also, that it is particularly hard to hit the spot in between - I don't know if he's correct about that.
I really question the idea that most people have any idea what their home is worth without relying on... Services like Zillow. Especially if they've lived there a while.
Similarly, if you think a house is worth $500k but you list it for $600k, you don't know if someone will decide they love that house and they're willing to pay that amount or not.
That's one reason Opendoor tends to list houses at a fair premium to what they paid, as sometimes someone decides it is worth it and they make 20%+ on homes like that.
If you ‘open up’ the flood gates on the other end, then yes you’ll do a lot of deal flow - Buyers sense a sucker - and open up a lot of opportunities for matches. It just so happens you’re also losing your shirt.
It’s easy to ‘make money’ (close deals) by giving money to people, and losing money in the actual business.
Are people interviewing all the neighbors before making a house purchase? Are those neighbors not incentivized to withhold any negative information about the neighborhood because it affects their own property values?
Indeed, I believe this is what OpenDoor does. From The Economist article [1],
"They [OpenDoor] charge a fee for the services they provide: buying and selling homes immediately, with zero fuss. The quick in-and-out makes them more like marketmakers than property investors, who buy to hold.
...
"A former Zillow employee told Business Insider that management had been hellbent on catching up with Opendoor, the front-runner. In order to compete, the employee alleged, the company pushed to offer generous deals to potential clients. It called this “Project Ketchup”. Now it has its own fake blood on its hands."
[1] https://www.economist.com/finance-and-economics/2021/11/13/a...
These are different things.
Archetypal market making involves simultaneously buying and selling an asset. Flipping involves buying, improving and later selling. One might be able to deal with the heterogeneity of houses by operating at scale. (Zillow attempted this.) One might also deal with the delay between buying and selling by hedging. (Zillow never seems to have thought about this.) But the improvement function makes what Zillow attempted fundamentally separate from market making.
They weren’t paid to provide liquidity. If anything, they paid a premium for scale and immediacy. They were a real estate operation masquerading as a tech outfit. WeWork in different stripes.
The main insights are that market makers hold assets for a short period of time making money on the spread between buyers and sellers offers. Zillow had to hold on to houses for a long time and was speculating that the houses would be worth more in the future which is not market making.
Does it? I worked for a few years for a market maker, and that's not what we did. Simultaneous buying and selling is what the arb guys did. We'd buy and sell with generally short hold times. Which makes sense to me given that the exchange has market makers to provide liquidity. If something can be simultaneously bought and sold, then the market-maker is unnecessary.
Archetypal, not predominant.
> Simultaneous buying and selling is what the arb guys did. We'd buy and sell with generally short hold times
The ideal market maker is arbitraging (and eliminating the arbitrage-able inefficiency). That’s why humans were replaced by faster-trading machines everywhere they could be. In most cases, the arbitrage is synthetic or approximate, e.g. hedging an options or swaps book. But a fundamental separation between speculating and marketing making is the latter does not take a view on the assets per se, and should not be betting on their future price movement.
No market maker always achieves the ideal. But they tend towards it. Zillow didn’t have that tendency. In fact, they erected fundamental obstacles between themselves and that ideal.
I’d have to dig up the textbook sources, but the key bit is in the definition: market makers quote a two-sided market and make money from the spread [1], i.e. buying at the bid and selling at the offer. If it happens simultaneously, that’s ideal. Every second one is long or short, risk and cost are incurred. Market makers seek to minimise and manage these.
In practice, arbitrage is tough. So most market makers simulate simultaneity by hedging. For example, if longs are accumulating (e.g. due to specialist obligations) one might open shorts or buy positional puts or wing it by shorting SPYs.
An unhedged market maker is just day trading.
[1] https://www.investopedia.com/terms/m/marketmaker.asp#what-is...
> 1) a statement, pattern of behavior, prototype, "first" form, or a main model that other statements, patterns of behavior, and objects copy, emulate, or "merge" into. Informal synonyms frequently used for this definition include "standard example," "basic example," and the longer-form "archetypal example;" mathematical archetypes often appear as "canonical examples."
> 2) the Platonic concept of pure form, believed to embody the fundamental characteristics of a thing.
The confusion between you two seems (to me at least) to fit almost entirely within the difference between those two definition. If you are describing the ideal market maker as essentially performing arbitrage, that seems to fit the second definition pretty well, right?
Meanwhile if wpietri says that most of the work at his believed-to-be-typical example of a market maker was doing non-arbitrage stuff, that'd make sense, right? I guess in most places the main work would be managing the divergence from idealness.
I think it also leaves out that not every market maker wants to be flat instantly. The one I worked for, and at least some of our peers were sometimes happy to hold inventory for a bit when they thought the market would even out.
By marking it up and/or appreciation. They buy it for $500/sqft and then rent it for $100/sqft. In that hypothetical the breakeven is 5 years, plus overhead.
If the occupancy doesn’t work out in their favor, they may still make it up in appreciation.
What’s a duration risk?
It reminds me of one of my favorite Bill Gates quotes:
"The first rule of any technology used in a business is that automation applied to an efficient operation will magnify the efficiency. The second is that automation applied to an inefficient operation will magnify the inefficiency."
Starting from scratch can be a huge advantage.
Yes, a poorly designed process sucks but it works at some level. That means the rough flow of it is figured out. Yes, there are exceptions and complications and all kinds of odd things but it's fundamentally different. It's not "from scratch" as you're taking an existing working-but-broken process where you know the input, know the output, and rethinking everything in between.
In an "inventing" scenario, you have what you think should be the input, a notion of what the output should be, and you're trying to build towards that notion.. without the validation that you're thinking of it correctly.
The first is a harder social problem (aka getting people to change) while the second is a harder technical problem.
But if you are trying to solve a novel problem, and the proposed solution involves "ML will magically predict the future", you'd better have a very good idea of exactly how the problems will be solved, or else you're probably better off starting with good old-fashioned human intelligence.
What often ends up happening is a large manual processes is automated bit by bit, and you end up with the situation you describe: a poorly designed manual process painstakingly replicated in code. Full automation is often never actually achieved here.
The absolute worst thing to do, though, is to begin automating the thing without fully understanding it. It's putting rocket boosters on your self-driving car without first understanding the rules of the road.
The techier folks definitely have a different set of problems but the speed at which hings get done is night and day. Companies with old school work patterns (which, in my personal experience, means dusty old banks) are terminally entrenched in their ways.
Taking some hopelessly byzantine, spreadsheet-driven process and “automating” it by building a Rube Goldberg scripting framework around it is the kind of totally stupid automation that doesn’t work.
Actually getting down to surface level and understanding fundamentally what each of those humans is accomplishing via those spreadsheets, extracting that all the way back out to a domain model and process flow diagram, and then selectively replacing process steps, whole cloth, with tech designed to be an actual subservice with SLA targets, is the right way to do it.
Throwing the spreadsheets and/or humans out altogether and starting “from scratch” is so exceedingly and needlessly risky from an information loss and hubris point that, well, good luck, but you’re nearly certain to fail.
Not to mention that generally ML models are not useful for assessing risk. ML nearly always focuses almost exclusively on some point estimate rather than a distribution of what you believe about a value. The former case is all about expectation and the latter about variance. Correctly modeling variance is far more essential to risk modeling than expectation alone.
I recall talking to a startup that was attempting to model credit risk by building a binary classier for defaulting, and trying to figure out a way to use this to score people for credit (obviously they chose to ignore the fact that there is a huge industry with decades of experience in assessing consumer credit risk).
They focused exclusively on finding more advanced models to get better AUC without even realizing that that's not important. I mentioned that the most simplistic credit score model should at least model P(default|info) and then set the interest rate to - P(default|X)/(P(default|X)-1) to break even and they couldn't comprehend this basic reasoning. It was doubly hilarious since their population's base default rate was such that the solution to this equation was higher than the legal limit they could charge for interest.
In the early part of the current startup/tech boom there was a focus on "disruption", the idea that new ideas could easily dominate old ways of doing things. But for many industries, such as credit/lending and real estate, you should at least understand the basic principles of how these "old ways" work before trying to disrupt them.
It is actually quite a common practice to design neural networks that output probability distributions.
This can be done for neural networks, through either bootsrap resampling of the training data or more formal bayesian neural networks, both of these are fairly computationally intensive and not typically done in practice.
This is the real problem.
Even if they have the historical data for that exact house/unit, it won't help them in cases such as:
* That nice view of the woods out the window is now blocked by a massive radio antenna that was just built there
* The river running through the back yard is now heavily polluted by something up-stream
* The new neighbor across the street is a huge nuisance and says they will never move
* The house just had a mass-murder event in it
Just because something is now cheaper than "comps" at price/sq ft and other metrics doesn't mean it's comparable.
Minor? Usually.
Random? Not at all. A minor annoyance like a cracked driveway ($1,500 to fix) is also likely to be associated with older kitchen appliances, faulty water pressure, deteriorating deck; poorly seated windows, etc. And then, buying that house for what the algo tells you -- or even algo minus 3% -- isn't likely to be a happy choice. Its fair market price may be algo minus 10% or worse.
Also worth bearing in mind, the Realtor community is not going to make life easy for Zillow. Once it's known that Zillow is loading up on clunkers, buyers' agents are likely to tell their customers: There's a Zillow house on the market, too. It's probably got problems. I'd demand a full inspection and some indemnities if I were you.
Common flaw of market disruptors. They assume that the existing players will remain neutral and indifferent to their arrival. The real world tends to be much tougher.
The opinion is famous not just for its unusual fact pattern but also because the Judge clearly had quite a lot of fun working in other-worldly puns and references while writing it.
This happens in hot real estate markets. If you don’t want to miss out or start a bidding war, you have to be the most frictionless buyer.
This entire topic is is marred with silly nonsense like "your only risk is mishandling cat feces", which totally ignores things like:
1) The cat litter box is covered in microscopic parasite-containing cat feces
2) The cat tracks feces-covered litter out of the box into the living space, which further spreads the parasite. If you walk barefoot through your house, you will step on cat liter pieces eventually which are again covered in microscopic feces.
3) The cat steps in litter covered in cat feces from other cats, gets it on its paws, which it then licks, starting the infection cycle over again. Multiple cats can easily keep a perpetual infection cycle going.
4) Cats, even well trained ones, routinely climb on top of furniture/eating/cooking surface and coat them in their feces. My parents were shocked after installing an interior camera when they saw their "well-behaved cats that never do that" walk all over their dinner table, kitchen counters, open and go inside cabinets etc etc. They are smart enough to know when you are around or not, and to not act like this when you can see them.
5) Cats have sharp claws, and even well behaved ones can accidentally scratch you or scratch a toddler that gets too close etc. This will then directly introduce cat feces to your blood stream.
Once exposed, the initial symptoms are mild flu-like symptoms and not a huge deal (unless you are pregnant, in which you will likely experience a miscarriage. Pregnant women should not be exposed to cats, or be in a household that has indoor cats). The real issue is the long term cysts they leave behind in your brain, causing a latent inflammation and immune response that seemingly interacts with your dopamine system in poorly understood ways. For example, latent cyst-stage toxoplasmosis significantly increases your risk of developing schizophrenia.
Originally the estimate on Zillow said my house was 20% over the value I actually sold my house for just last month. I listed with a traditional realtor for a 5% commission, because when I looked up the service and other fees for Zillow sales, I found they included around 20% of cost for buying homes and closing within generally 10 days.
As I listed my house, and as I reduced price on it for it to gain attention, I noticed the zillow estimate also went down to always stay below my listed price. I believe the estimate that both Zillow and Redfin display prominently were purely based on what my list price was changed to last, not on any meaningful algorithm, which can be very harmful to sellers and buyers, because it makes the process a bit deceptive by nature. Luckily Zillow also displays the price history on homes, which apparently cannot be "gamed" as much as the "zestimate" can be. Another thing I noticed was that the view stats on my listing that zillow regularly provided changed, even after days passed, that was very concerning because stats of that kind aren't supposed to change... They indicate real interest in a property, that guide decisions for sellers to reduce price, and they also indicate what is truly a "hot home".
No matter what, there is always the "human factor" that can corrupt or even destroy any company, where realtors can game the process to maximize their own sales profit or positions, or where appraisers can inflate an estimate as a favor for a personal friend, even despite laws against doing so. In a bad economy, the lengths people will go to to suit their advantage are wild. This type of issue can never be properly addressed by any algorithm, and that's why trusting technology too much can so easily lead to failure in any setting.
Ultimately I am glad I did not sell to Zillow, because of all of the potential for hidden costs and because they manipulate the process even when you don't use their service, but I am not feeling sorry for them as a company... I felt the impact of their presence in the market whether I involved them or not, and that's a big problem when it comes to preserving the value of traditional investment and stable investment in a house that should be properly addressed by regulation.
The fact that a house is for sale at a given price, but has not sold after some time, is a strong signal that it's overpriced. The longer it's been sitting, the stronger that signal is. They'd be crazy not to include that data in the Zestimate.
Now, if it's extremely fast, eg they adjust the price down within a day or so, then it seems a little ridiculous. OTOH the Zestimate has always been a rough indicator at best.
I am luckily both a web developer and knowledgeable about real estate, most people don't properly understand the dynamics that are impacted/introduced into the market by technology and algorithms... People assume the traditional real estate market rules are still in play primarily still, but technology has complicated everything... That's also why Zillow overbought homes, because people making critical decisions too often put "traditional pre-tech" real estate market concerns over considering modern impacts of IT to their decisions.
My house sold within 2.5 months overall, it was not on the market for a long time.
I wish I had the data that I assume they have internally, because watching their actions I’m not convinced they understand what questions would actually be interesting to explore with ml.
Why would Zillow have unique insights? With the exception of Texas, I thought real estate sales information is public information in the US.
How many people search on bedrooms but not bathrooms? When people search on both, what’s the pattern they use? If we highlight prices and BRs on the map does that give more clicks than just prices? How important are photos (times 50 different questions there)? How strong a signal is repeat views spaced over time? Saving a house to favorites? Sending a link to a friend? Clicking on comps in the neighborhood? Which comps do people zero in on (as evidenced by spending more time on the page)? How strong a signal is sending a message to the real estate agent on the listing? What areas of the country are seeing an uptick in search traffic? How long between claiming a house as an owner on the site, updating the information, and listing it for sale?
They are sitting on a (well-earned) treasure trove of data and it’s not unreasonable to think they could use that to be better informed than another buyer without that information.
Where they seem to have failed is in not augmenting that advantageous data with regular old boots-on-the-ground observations.
It is very difficult to go from what users browse to what they actually buy. People very often say one thing, then do something completely different.
And sometimes they browse stuff just to make sure that their current decision is correct, so they will look at a lot of items they're not going to buy.
(oh, and everybody and their mother knows photos are important. No need for ML to find that out)
So it is easy to test UI changes, but difficult to find out why people do what the do.
Click data is much less valuable that the recent sale price data available in MLS. Using 90s style dwell time and click counts would likely yeild a lot of very noisy data. False positives from people's browser reopening with 15 tabs looking at different houses. False positives from social and paid advertising boosting a particular home or neighborhood's numbers. False positives from enterprising real estate entrepreneurs doing everything they can to get the clicks up in areas they own property to drive up prices. Meanwhile, the recent sale prices tell you much more, with certainty and are very expensive to manipulate.
Also, unless Zillow started imposing confidentiality agreements on their bids, then competing buyers would just have to bid $1 more without their dataset, right?
In my mind this is the problem with consultants who try to automate processes. It’s really difficult (maybe even impossible?) to successfully write a program to make a computer do $thing if you don’t understand the intricacies of how to do $thing manually.
Houses trade slowly, so would sit on Zillows books for a long time (days/months). Market makers on the stock market can have assets sit on the books for under a second. Houses are not fungible, which extenuates the slow trade problem.
At a guess, in our county, 20%+ of the housing is idle, owned by out-of-state companies, some of whom pay property taxes and some dont. The county isn't auctioning off because of tax default anymore, no one was buying these places at $100. Many of these places are complete teardowns now; some actually no longer exist, having burned or apparently been scrapped. The tax assessments on those have not been adjusted, for the few i checked.
I think the housing market is so fucked no one really grasps the scale of the problem.
I don't think I agree with this assessment. I live in a very rural area two hours northwest of Austin, literally in the middle of nowhere. I've studied the local economy and understand how things work here.
I think the characteristics you've identified in the rural housing supply are not unusual and also not as serious in a practical sense as you seem to be indicating. For example, in San Saba, Texas, 20-30% of the households are under the federal poverty threshold. The median household income in the town of San Saba is about $32K/yr. People just don't have any excess cash so the maintenance on dwellings is neglected. That means folks become extremely thrifty and resourceful patching what needs to be patched, very cheaply, if not for free. Some dwellings simply aren't maintained and one day won't be there anymore.
Families live on small budgets, don't require much and generally just "get by". The municipal and county governments have very small budgets but extremely resourceful staff who accomplish a lot with very little. Everyone comes together as a community when needed (see: February 2021 freeze event) and it all works very efficiently, actually.
To someone who is not from here and who doesn't understand that dynamic, they might see those properties as you described and believe a tragedy was unfolding. But that doesn't reflect reality on the ground vis-a-vis my neighbors.
Our local Craigslists always have "Property inspector" jobs listed. You go take some cell phone shots of buildings to prove they exist. The people I have spoken to who have done those say they didn't bother going to the places as often as not and took pics of some neighbors house. Even when people actually do that job and document the true state of these properties I can't help but suspect the information is buried or lost because thats not the narrative management would want.
The actual family owned housing stock got better the last two years, our population doubled for the last 3/4ths of 2020, and all those relatives did a lot of renovation and rebuilding.
I understand what you're saying. The ripple effect created by that dynamic would unjustifiably inflate local property values, reducing affordability for locals, creating synthetic demand by reducing supply as the land could otherwise be auctioned.
and (ahem) East TN is more "western Arlington VA" IMO. I said rural. I'd have to walk a half mile to get a decent rifle shot at a neighbor. It's getting too crowded here.
But there are so many abandoned places out here. People have just walked away and never looked back. We had one across the road that over the last 10 years the woods has reclaimed and unless you knew that it was there, you would drive right past it.
In your opinion, what do you think the most effective way to help these families out might be?
This is a question that I'm well-positioned to answer. I moved to this rural area in 2018 after living in Austin for 24 years. I immediately looked for ways to volunteer and help.
I developed relationships with elected and community leaders, started my own "technology incubator" to teach technology skills and classes. I explored establishing a regional technology council with my county judge and Texas state leadership. The community liked that I was volunteering but the actual uptake, expending effort to learn and implement what I was teaching, wasn't there. They didn't know what to do with it. The gap between their world and the world we know at HN was too wide to be bridged effectively.
My experience is applicable to every problem here where someone thinks they may be able to help in some way. Whether it's teaching job skills, helping those who are addicted to meth or whatever, I believe people can't be helped if they don't want to expend the effort to get from A to B themselves.
There are many reasons for this, why offering to help in an economically-depressed or disadvantaged community doesn't yield results. Locals are apathetic, comfortable living in the middle of nowhere with very low expectations, or else they have poor self-esteem and don't believe they can do better.
I don't "push" anymore. I just try to be empathetic and understand their situations. This past Thanksgiving I asked the community to tell me if anyone was unable to get a turkey for Thanksgiving and would like one. Two families responded; I was glad to help. It's little things like that which I can do to help their situation which I feel is the best approach now.
Edited to add: There is an organization here called "Mission San Saba" where a group of ~30 volunteers will pick one house per year to renovate, typically for an older or economically-disadvantaged family. That has been very successful here.
San Saba ISD is probably the best funded entity in the whole county. Every student has a laptop and home internet. The graduation rate is 100%. It's a small school; the senior class is only 50 students.
They built the new school in the middle of town, thus highlighting its position of import within the community.
Washington state does something similar, though it’s more of a subsidy. Education is still mainly funded locally, but the state kicks in with its own funding for poorer districts, so Seattle property taxes subsidize schools across the state in Spokane.
Arizona legalized medical marijuana quite early, followed by recreational marijuana. Their medicaid program AHCCCS [1] is extremely comprehensive and even pays for Uber/Lyft to the doctor's office and back. Patients are able to see a great selection of GPs and specialists, and the copay is always $0. The accompanying drug plan is comprehensive, also with a copay of $0. AHCCCS will approve expensive modern drugs like Rozeram (supercharged melatonin analog for sleep) if the sufficient documentation of reasonable need is provided.
Cactuses are protected from destruction by law, and must be transplanted when doing clearing for construction. You may find the idea of being able to own a firearm without a license to be unpalatable but the state largely remains very safe crime-wise (perhaps due to that?)
I miss living in Arizona. It's a beautiful state with very caring folk. I saw almost no homeless folks in Phoenix. Folks there seem to really care about their fellow citizens. Southern hospitality is for sure a thing, take it from a daft boy from Brooklyn!
It sounds like that funding decision predates the current crop of state leadership
I've seen this personally, too. A house I rented until a couple of years ago was owned by a Chinese company, which also owned half of the other houses on the block. We all paid rent to the same LLC that forwarded the cash overseas, and did almost zero maintenance.
I think the housing market is so fucked no one really grasps the scale of the problem.
One thing I don't see discusses very often is the affect that large "master-planned communities" have on a city's housing prices. I've seen at least three cities where mega developers like Howard Hughes Corp own massive tracts of land, but instead of building houses, sit on that land waiting for the price of housing to go up. Sometimes the developers are very open about it. Sometimes not. But instead of allowing a free market to develop 5,000 new homes, they develop one lot here and one lot there.
Or worse — I've seen them build hundreds of homes and then sit on them, empty and vacant, waiting for prices to climb high enough to put the houses on the market. Again, a drip at a time, to keep the housing supply artificially small so they can boost their profits. Meanwhile, people have nowhere to live.
1. The property must be residential 2. If the property is commercial, a foreigner must incorporate in China 3. You may own only one property 4. You must possess a long-term visa 5. You cannot be a landlord, as a foreigner 6. You must pay a 1% deposit and an initial 30% of the purchase price to the seller in RMB, if you are obtaining a mortgage
I think there needs to be a campaign to educate the US population on how lopsided the system is. Only then will politicians start to get massive heat to enact at least equivalent restrictions. Right now its all the rich and foreigners picking away at the carcass that is America.
What's the issue with out of state companies owning rural properties nobody wants? If the market is heating up, maybe it's time to run tax auctions again.
In WA state, if there's no bidders, the county retains the land and will auction it again when someone expresses interest (or it some cases, can sell it to a neighboring land holder without auction, like for the 1930s era tax foreclosure I bought last year)
Jobs generate demand for homes. Homes cost a lot in areas where there's been more jobs added than homes. In the last decade, the bay area has added seven jobs per every unit of housing constructed.
The amount of social media content revolving around "how I became a milionaire/how I reached my first million" and the common factor is "I bought a house in 201*", then I'd say something is a bit off...
Either there's massive speculation, or 1 million isn't what it used to be, or worst: both.
The problem is that their blogging about it attracts the people that want to get rich quick and they are the ones likely to lose their shirts.
If this kind of algorithmic speculation took off, not only are we likely to see the algorithms themselves form feedback loops to push prices up far higher than otherwise, even if they don't sell directly to each other, I believe it will drastically increase the size and rate of the boom bust cycle.. and we're doing that to _peoples homes_.
The scope of human suffering possible here is huge, the societal damage massive.
This needs to be so very very illegal.
i think zillow also did not understand what fungibility means. real estate is not fungible. in fact, it's anti-fungible- that's why there are huge diligence processes that exist around most real estate transactions. maybe floors in an office building may be fungible, but residences are definitely not- with all their quirks, customizations and problems.
this whole argument that they failed because the ceo of zillow didn't have big balls is pretty putrid. pair this with the word salad of misused words and twisting of quotes and i'd say this is probably one of the worst pieces i've ever seen posted here. i feel worse for having given it any time at all.
the simple fact is that the ceo of zillow didn't know what they were doing, had a team that (supposedly) applied facebook's infrastructure scaling prediction library to house prices and then attempted to apply a market making mindset to the real estate market at scale. not only is this probably something nobody should try to do, considering we contribute so much in tax dollars to first time homebuying incentives as it's recognized that the housing market is where j q public can start to build wealth, it's also probably something that nobody could do (well, at scale, with machines) given that housing is not fungible.
sure, financial instruments are fungible, maybe even late model cars, but definitely not houses. doesn't take a scientist to spot that.
If buyers want more now for their house, than it can be sold for in a few months time (which is necessary for renovations and other prep for sale), then there is no ML (and no non-ML) method to make money. Either you overpay and lose money, or you don't overpay and you don't buy any houses.
In that situation, the only smart play, is to get out of the market. Zillow is, no doubt, not perfect. But they have a lot of knowledge of the housing market, and they thought it was time to get out entirely. I think the author of the article either isn't able, or doesn't want, to consider that Zillow might have been exactly correct in doing so.
1) Why layoff your data science division if they are predicting with accuracy?
2a) If you have enough conviction to call the top of the market, why sell off so much housing at a huge loss? Zillow are the only participant in the residential real estate market losing money right now.
2b) If you see signals of a forthcoming housing crash, why not short the housing market?
The simplest explanation is that Zillow was poorly run.
- they could not buy houses without overpaying (relative to what they could sell them for a few months down the line)
- the housing market would not recover for several years (so no need to keep that extra 25% of your labor force, especially if you anticipate a decline in revenue from real estate agents coming soon)
The issue is that the housing market is just unsuitable for this strategy. Houses aren’t fungible, and they are very slow to trade. So Zillow ended up in a position where rather than clipping the ticket on spread, they were actually quite exposed to house price movements.
This kind of thing is difficult to confirm from the outside, of course. But that they adjusted the model to pay more towards the end is pretty widely known.
https://ryxcommar.com/2021/11/06/zillow-prophet-time-series-...
"Speaking of middle managers, word on the street is that Zillow Offers put their thumb on the scale of the algorithm to make it engage in more aggressive trades. Manually adjusting an algorithm isn’t necessarily a bad thing, but you need to do it for the right reasons. And clearly that didn’t end up working out..."
That not everyone in the organization knew that, is just me speculating.
Actually that piece also hints at an uneven understanding of these issues across the organization:
> In one tweet, I semi-joked that what happened was lower and mid level employees convinced upper management that the algorithm was 99% accurate by hiding the caveats of what “99% accuracy” means.
Vast swaths of rural Midwest and northeast with little industry and declining population definitely did not make out, especially factoring in property taxes and the opportunity cost of not investing in VOO as a near risk free alternative.
The most valuable data is not social data, ... but your own data because every dataset that you’re looking at internally describes your own process, including your bugs, ... building models from your own data is the only way to build a really successful system.
This is one thing that a lot of outsiders do not understand. Facebook/Google's data is basically worthless to anybody but Facebook/Google. The data has value because it is derived from their own processes, which in this case are the requests and context of each product surface.Facebook and Google's data are not their own. That data is comprised of private lives, stripped bare pixel by pixel, bit by bit, and it's offensive to frame it as if they're doing something alchemical and special with it. Google's search dominance came from something special, creating the right algorithm and seizing the first mover advantage, but the relentless and ruthless invasion of privacy is a rent seeking race to the bottom.
All of the ills of the internet and political turmoil in the west from algorithmic amplification are the brainchilren of Facebook and Google. It turns out that "tailoring search results" and "targeted advertisement" are excuses for something that can cost far more than a society might want to pay.
Most data Facebook collects is of the form (user saw this post, user clicked/did not click this post). That data's value is tightly coupled to the process Facebook used to decide whether or not to cause the user to see that post. The data only has value in the context of iterating on that process.
You’re also absolutely right that the social media content: the photos, the sentiments, the likes, the connections, should not in any way “belong” to FB/G.
The data that does belong to them, and that is useless to anyone else, are the outputs from their sentiment analyzer service, the weights and trigger conditions for their content ranking algorithms, the intermediate outputs of their ML evaluations, etc.
GP, and the article, are saying: look there first. Try to start by truly understanding “what you already know, but aren’t paying enough attention to,” and don’t just treat the problem as “needs more data.”
> rent seeking
> first mover advantage
This reads like an HN buzzword bingo card.
But seriously, a lot of your claims are flimsy or misinformed, which goes to credibility. Cambridge Analytica was a huge nothingburger that had no actual effect on US elections or Brexit. Google did not have first mover advantage, they were so late to the search engine game that it caused them trouble in their early financing. Show us a real, known harm from Clearview. Google and Facebook are not breaking any laws, so how can they be "invading privacy"? The bottom line is, people love FAANG tech, and are happy to trade their data to use it. And one of the reasons is because they are not experiencing real harm, in spite of what HN's white knights would have us believe.
Some domains are intricately mapped in available data (e.g. equity pricing), but most, and especially most physical, are not (e.g. freight transportation).
In particular when you’re auto-underwriting credit it’s not typically an origination-for-sale model. So the value of the loan is the present value of the future payments, less the future value of defaults, less the cost of acquiring the customer.
Historically those things can be modeled pretty accurately and the aspects that can’t be modeled accurately can often be hedged or eliminated by the law of large numbers. The innovation of the new ML underwriting with respect to accuracy is at the margins. The real disruption is the speed and cost. (Disclosure: I worked at a SMB fin tech and we reran multiple credit models for a million customers and past customers every night.)
If Zillow were getting into the rental business, in some ways it might have been easier for them. But they needed to model where they could sell an illiquid asset which is a much harder and much less well understood problem. And yes with enough capital to plow through and the appropriate risk attitude they could likely have gotten the handle on what their pipeline was really going to look like. But it’s hardly the same problem as credit underwriting.
Good tweetstorms with technical explanations on how that happened:
https://twitter.com/macrocephalopod/status/14558873523715973...
Ultimately they were really bad as flippers. More often than not paying more than market price for the homes they bought.
I think the root problem is that this was a panic move. They saw Open Door's success and thought they had no choice but to try and replicate it. But its a questionable business move for Zillow and ultimately they couldn't make it work
Zillow realized the only time their ask was hit is when it was at a premium to the actual market price. If they used competitive offers, they’d never have the winning bid. In a hot market where you’re offering a premium, you’re going to have owners of lower quality properties accepting your offer, while owners of higher quality properties have more offers to select from.
Zillow got left holding a bag of lemons and decided to get out before buying the whole lemon grove.
Why do you assume that, seems like a cash buyout would be a great deal for many sellers if it was at the appropriate price. Issue is I think that Zillow's information was less granular than what the buyers/sellers had. Let's say Zillow priced two houses near each other at 1million each. However one was close to a busy road so would only sell for $900k while the other could sell for $1.1. Zillow made the right average offer of $1million to both but the buyers/sellers actually had more information. So the 1.1m seller didn't take Zillow's offer while the 900k seller did. Now Zillow was out $100k essentially not counting fees.
The problem seems more that they were not getting “enough” houses doing it this way, especially competing against Opendoor, and so they had to bid higher and on more properties in order to hit “scale”. And that lack of selectivity is what led to the bad basket of houses they now own.
The issue is that their machine learning model can't possibly be 100% accurate, there will be some amount of error that is shaped in a normal curve.
If their model overestimates the market value, they end up massively overshooting their goal price of "slightly less than market value", the seller accepts and they lose money. If their model underestimates the market value, they will offer way too little and the seller will go elsewhere.
Even if they get their estimates right 99% of the time, the 1% of cases where they get it wrong will slowly drain money out of the scheme.
Of course iBuyers can’t perfectly forecast the market but that is why they add 3-7% fees, a very large buffer on a house purchase.
Again, this is where Zillow ran into problems: they reduced or eliminated that fee to win more deals versus opendoor.
Planning to lose money takes nerve. Zillow tried to avoid avoid the pain, and ended up abandoning what might be a profitable enterprise (for someone else) in the future.
That's a fair point; the essay doesn't do much to distinguish whether they didn't know they needed to take losses, or couldn't take the pain of the losses.
Nevertheless, it's a pretty good analysis of what a company needs to do, in order to build a model relevant to their own actual business. They need to both know about the pain involved, and be prepared to take it. (And even then it might not work!) Third-party data (and suffering) might not be a good substitute.
Risk aversion and launching a new business strategy do not work well together.
CEO said cut! Way to go!
This loss was not immaterial but it also wasnt too material as they werent even leveraged on the homes. They had orders of magnitude more capital to risk if they really chose to dive into this or take it at least to real estate 2008 levels. Far from it.
This is similar to 'adverse selection' in real life & in Zillow's model. The article makes a nod to this, but seems to imply that if you train your model on that adverse selection, you can come out ahead after paying to learn about it.
To me that kind of misses the point. Adverse Selection isn't a static feature of the landscape you can identify and avoid, it is people understanding what you understand, adapting, and responding. Train your model with adversaries trying to beat it, then you'll maybe counter the specific first round strategies they use, and they'll learn new ones and beat your new model with their 2nd round strategies. It's a continuous game. Your requirement to gather a corpus of training data will keep you in the 2nd turn of a game where the wins are biased to whoever has the 1st move.
Cryptocurrency is also a bit wonky because of always including forever lost access to a solid percentage of the currency. Bitcoin is the most notable.
The thing with crypto is that much like some of these other commodity markets there's less real trading volume than many people think (there's been a lot of wash trading going on: https://cryptobriefing.com/binance-wash-trading-icebergs-tip...). Where crypto is very different from the futures markets is that you can just buy the stuff directly because the costs of holding it are much lower. Say I want to invest in oil, it's a massive pain in the ass to build warehousing to start taking delivery, whereas something like crypto is much easier for a company or individual to hold. From this point of view there's very real non-regulatory reasons why trading futures for oil makes sense whereas this is not so for cryptocurrencies.
Sure they failed. But the only data we have is that they failed because of something very specific which doesn’t relate to much else.
This is what most profit seeking strategies can miss. Their designers (consciously or not) can't help but to stop thinking through their plan at the profit step and just assume "rinse and repeat" forever after.
Opendoor's primary benefit is to enable people to move when they otherwise could not easily do so, creating more liquidity and matching supply and demand (often number of bedrooms in house to number of bedrooms now needed).
The challenge with moving is that most people need to sell their current house before they can afford (or even know what they can afford) to buy their next home. Opendoor lets a family buy that next home with its cash, then list their current home on the market or sell it to the company so they avoid the double mortgage or double move (home->rental->home)
In contrast, “Zillow Seeks to Sell 7,000 Homes for $2.8 Billion” so Zillow lost more than a few percentage points.
It's a hybrid model trading in an adversarial, real-dollar environment. The leverage comes from having a small human team trade big volume, much more than they could possibly trade directly, by augmenting their human abilities with automation and a model. Or seen from the other side, it's a model with human oversight.
Any system like that is high risk, high reward. All the successful ones started out by losing a lot of money. Paypal lost an incredible amount to fraud before they started breaking even. OpenDoor lost an incredible amount to mispricing, and took on a ton of balance sheet risk, before their business really started working.
"To live, you must be willing to die"
- poker legend Amir Vahedi
A big part of Opendoor is creating the right apps and processes to collect this information to feed their models. The machine learning part is important, but can give the false impression it's just about data scientists crunching numbers at head office, when in reality there's a huge real-world operational machine that's driving it.
I see this as a victory for us calling out companies for immoral behavior.
It's telling that the article opens with a clip of Alec Baldwin talking about needing "brass balls" from Glengarry Glen Ross, seemingly oblivious to the fact that it's a dark comedy mocking cutthroat sales culture.
This article does not mention that. Instead, the rest of the article deals with Linkedin-wisdom and hard platitudes, such that it is not possible to build a good model on someone else's data (as if Zillow even was).
Data scientists remarking on the Zillow fold, are like psychiatrists or engineers remarking on non-clients and bridges build by others. They know nothing about the business, about the constraints, about how the estimates are consumed. They end up silly, but without good information coming from Zillow, we assign value to their analysis, purely on Twitter-soundbite-ability and internet-authority.
During the time of the pandemic, the house prices rose and so did the volume. If anything, they made out like a bandit.
False. You can definitely bootstrap and adjust the model as you either gather more data yourself or get more outside data. You can also build confidence intervals around the model predictions and decide how you want to proceed based on that. There is lots you can do with that initial model.
Many seasoned wall street algorithms have suffered many times over 5 decades, and when they fail we call them black swan events.
The author is arguing that they should have pivoted from “we already have models” to “we’re intentionally gambling hundreds of millions of dollars so we can build good models over the next few years”. That might be a good strategy for a startup with loads of VC money and no other products, but it makes less sense for a more established company to risk going under on that bet
He uses Zillow to explain how datasets – especially the ones with money tied-in – can’t be trusted blindly. Building a high-quality dataset is an expensive endeavour.
Not in my book. All I see is the price of real estate being driven up by corporate greed and the individual home-buyer being shut out of the market.
Is it wrong of me to hate "flippers" (be they corporate or private)? Pure capitalists will tell me that every property sold went to the highest bidder — in the case of a flipper winning they were willing (able) to risk the capital to hopefully turn a profit on the flip.
I suspect if you dig deeper you might find sales going to flippers because they had 100% cash offers, because they are better at "the game". I see no reason to punish prospective first-time home owners in this sort of market.
But I don't know what the answer is either.
If houses were a (much) smaller bet for the buyer, there would be more flexibility to build houses where demand exists and a faster, lower-drama exit for people who don't like the changing nature of their in-demand neighborhood.
The inertia created when people have their life savings tied up in their house perpetuates the problem of affordability, by making the areas that have the most mismatched supply vs demand the least likely to deal with the problem.
The process of building new structures is filled with so much regulatory friction that it is impossible for the average person to even consider building their own home.
Developers will always attempt to skimp on quality to save/make more money. Even people building their own home will sometimes try to avoid compliance. That's why the regulations are there.
Single-family zoning is another local government policy that is absolutely intended to constrain development, not improve safety.
I am in favor of finding ways to encourage more housing, but what you're calling for is essentially to invite favela housing in the developed world.
Parking requirements are about local traffic management as well. Set backs are about ensuring natural light. Some local regulation is about NIMBYism or HOAism, that sort of thing is where reform might be better addressed.
That housing should also be up to a similar standard in terms of its externalities like pollution and energy efficiency etc.
We have regulations for air travel, for car emissions and efficiency, why should housing be any different?
We have regulations for car emissions, and that raises the price for cars, which means some people can't afford cars.
We have regulations for housing, and that raises the price for housing, which means some people can't afford housing.
How many people should not be able to afford housing? Is the number of people who currently can't afford housing too low, or too high? Should we increase regulations for housing, or decrease them? Are we making the right trade-offs?
And this is accomplished via building codes, which are rigorous and applied almost uniformly in the U.S.
> We invest (or should) a lot of our taxes into local amenities to ensure that housing is provided the best environment. Transport, schooling, roads, etc.
And this is the model that has made blue cities unaffordable for the poor. They're not environmentally friendly either, their schools are awful, amenities poor, and transportation lacking. It would be hard to find one single issue where there is even parity of centrally planned quality-of-life concerns in blue cities vs red cities.
The question is not "regulations" persay. There is no magical regulation slider bar that can be adjusted to optimal result. It's what those regulations seek to accomplish. In many U.S. urban metros, those regulations are targeted to what city policy thinks the owners should do with their property, and not what they want to do with it. It's not clear those regulations have had their intended effect.
All successful work probably displaces someone else in some way. If you're good at your job, you're "denying" that job to someone less skilled. If you work in software, you're automating things that would require more labor if done manually. Fortunately, humans can pivot.
Either hate everyone, or hate no-one. You can't just hate flippers.
House flipping isn't a social negative. They're doing a productive activity and producing value. They aren't long-term speculators removing housing stock from the market. It's essentially home renovation, done by a 3rd party owner.
It's like Trump voters who criticize "libtards." We all know it's not a valid way to discuss something.
In fact, the expansion of the fiat money supply enriches the wealthy through the Cantillon effect. Then, because the value of money is going down, they pile into assets like housing. For instance, the US is becoming a nation of renters due to this effect. The stock market is similarly distorted. We need bitcoin because we need an objective form of money. That would allow stocks and houses to stop being stores of wealth and reflect their true economic value, which would be a huge boon to everybody.
I'm guessing the energy thing is what you think your anti-bitcoin argument would be. Bitcoin mining is also such an efficient market that in the long run, only the most efficient forms of energy--such as nuclear and geothermal--will be viable for it. Bitcoin is already helping to advance "green" energy. This is abundantly clear to people involved in the mining industry.
Who is doing the selling? "Wall St Fat cat Co" or the average Joe who saw his house value go up by a LOT?
When average homeowner Joe sells their house they still have to live somewhere. They must immediately use that money for another house, which is also inflated. The higher sale price doesn’t matter.
Average non-homeowner Joe trying to buy a first house is SOL.
Flippers take the risk of the market falling while they're flipping - that's the price they pay for their profits.
It pains me to watch people apply simplistic theoretical laws of supply and demand to something as complicated as housing. The map is not the territory. There are massive costs to increasing supply, as well as psychological/community costs to moving homes, which are not cleanly captured in any Economics 101 textbook.
I think you’ve bought into the tik tok narrative that somehow it’s zillows fault that houses are expensive.
I'm buying a long term asset, so the liquidity of the housing market is not relevant to me, unless I'm actually buying for a specific short term, like a planned work period.
Liquidity of the housing market is only important to the agents and the loan originators because they make money on the flow.
No. Like a normal person, I bought my house to live in and to improve and to stay in for a long period of time. It isn't a speculative investment vehicle.
It would not change my physiological need for shelter, no
“Oh I might not be able to sell this for a profit in two years, guess I’ll die in the street”
As a seller, would you rather wait a year to make a bit more money? That wouldn't be good. That would be crappy.
> A machine learning organization thinks of risk entirely differently than an automated risk underwriting organization.
It's possible and maybe even advisable to use machine learning in the automated risk underwriting business, but it is a different setup / set of objectives.
As the author notes, IMO the adversarial and antifraud aspect of risk underwriting turns it less into a straight-up estimation problem and much more into a game theory type of problem. ML models can assist in evaluating risk, but you do indeed have to be preocuppied by your risk as a party to the transaction in the first place, and not just trying to predict prices as a third party observer (which by itself is pretty riskless).
And if you have billions of dollars in cheap capital, everything looks like an investment problem.
Which is ultimately the suggestion of this article: "Why aren't you more like Wall Street?"
The implications are exactly the opposite of Zillow being an innovative company. If they require billions of dollars in deep pockets (nbd) and a restructuring of their org to be more like old-school operators, all signs point to existing players as more fundamentally correct about the strategy required to succeed in the space.
If Zillow thought they had all the data they needed, there would have been little harm starting with $100 million in properties -- if the loss there ended up being $5 million, they would have known immediately something was up and that they had work to do.
This is ridiculous, we need much better regulation on this stuff.
I wonder if higher property taxes would help a bit? If you own a 'home' then you're going to be paying for the water, school, electricity infrastructure whether you use electricity, water, or not.
Of course, that would be gamed hard and would have to be strongly regulated as well.
But that, and vacant property taxes, limits on some other things, and some other adjustments might help.
You're not going to be able to reliably model asset prices at the resolution and accuracy needed to front run the market for a long period of time. This case was worse because the "Zestimate" directly created a feedback loop that moved the underlying asset prices higher.
Most AIs today are for augmentation, not replacement. Vehicle autopilots are a perfect example. The ones that are commercially available aren't capable of replacing the human, they just augment the human's abilities.
No pricing model will ever get this right.
Always has been that way. Always will be that way. AI is great for when you need to tame a firehose and make millisecond decisions. But there's a 90 year old in Omaha who is better than the best AI.
What does this mean? 50% of the money is held as debt? Or 50% of the money is lost to fraud?
One example: The MLS in Austin, TX recently banned publicly sharing a home's sold price. https://www.zillow.com/austin-tx-78701/sold/
How ?
(A) Both Wall Street and Machine Learning Modelers struggle with tail risk. Hedge funds measure performance against
https://en.wikipedia.org/wiki/Sharpe_ratio
which assumes risk is (i) normally distributed and (ii) a source of reward. For most people, however, risk looks like Theranos or the Fukushima accident or the Challenger distaster.
It's unbelievable that a machine learning model trained to predict house prices based on experience would be accurate in the face of events like the COVID-19 pandemic or what will happen when the Fed raises interest rates. You can model risks like that, but to the extent that you're working from experience you are working from a database from the 1929 Crash, South Sea Bubble, etc.
(B) Mark Levine wrote a good article about how you'd exploit such a predictive model. If you consistently gave people low offers, a few people would accept them. You would get a high rate of return but could invest little capital.
To invest more capital you have to make more offers that get accepted, that is, give better prices. Your rate of return goes down and if there is shrinkage from errors, accidents, etc. you could get a negative return.
It's that "tendency towards a declining rate of profit" that Marx warned about.
(C) The analogy with stock market market makers doesn't sound good when you consider the differing timescales.
Market makers are isolated from some risk because of the length of their holdings. Yet, they make profits by exploiting the stochastics of a stationary market (e.g. if you don't like the price at time t1, you will usually get a better price at t2) but they lose money when markets move definitively in one direction or another.
That kind of trader heads for the bathroom when things go South and in the interest of being orderly markets impose sanctions on market makers who do the natural thing and press the "STOP & UNWIND ALL POSITIONS" button when it gets tough.
In the case of Zillow I see holding times that go on for weeks or months and all kinds of real world risk like planning to do certain renovations but having to delay the work because out of 20 things you need from Home Depot they only have 16 of them.