Stable Diffusion 3
stability.ai
stability.ai
Some notes:
- This uses a new type of diffusion transformer (similar to Sora) combined with flow matching and other improvements.
- This takes advantage of transformer improvements & can not only scale further but accept multimodal inputs..
- Will be released open, the preview is to improve its quality & safety just like og stable diffusion
- It will launch with full ecosystem of tools
- It's a new base taking advantage of latest hardware & comes in all sizes
- Enables video, 3D & more..
- Need moar GPUs..
- More technical details soon
>Can we create videos similar like sora
Given enough GPUs and good data yes.
>How does it perform on 3090, 4090 or less? Are us mere mortals gonna be able to have fun with it ?
Its in sizes from 800m to 8b parameters now, will be all sizes for all sorts of edge to giant GPU deployment.
(adding some later replies)
>awesome. I assume these aren't heavily cherry picked seeds?
No this is all one generation. With DPO, refinement, further improvement should get better.
>Do you have any solves coming for driving coherency and consistency across image generations? For example, putting the same dog in another scene?
yeah see @Scenario_gg's great work with IP adapters for example. Our team builds ComfyUI so you can expect some really great stuff around this...
>Dall-e often doesn’t even understand negation, let alone complex spatial relations in combination with color assignments to objects.
Imagine the new version will. DALLE and MJ are also pipelines, you can pretty much do anything accurately with pipelines now.
>Nice. Is it an open-source / open-parameters / open-data model?
Like prior SD models it will be open source/parameters after the feedback and improvement phase. We are open data for our LMs but not other modalities.
>Cool!!! What do you mean by good data? Can it directly output videos?
If we trained it on video yes, it is very much like the arch of sora.
Very interesting. I've been streching my 12GB 3060 as far as I can; it's exciting that smaller hardware is still usable even with modern improvements.
Bigger than that is also possible, not saturated yet but need more GPUs.
Why is there not a greater focus on quantization to optimize model performance, given the evident need for more GPU resources?
Need moar GPUs to do a video version of this model similar to Sora now they have proved that Diffusion Transformers can scale with latent patches (see stablevideo.com and our work on that model, currently best open video model).
We have 1/100th of the resources of OpenAI and 1/1000th of Google etc.
So we focus on great algorithms and community.
But now we need those GPUs.
It’s StabilityAI that makes Stable Diffusion X.
2. the old advice is to sell shovels during a gold rush
Historically, once the easy to find gold was all gone it was the people who owned the deep gold mines and had the capital to exploit them who became wealthy.
The reason is because they instantly get a risk free guaranteed VERY healthy margin on every card they sell, and there's endless customers lined up for them.
If they kept the cards, they give up the opportunity to make those margins, and instead take the risk that they'll develop a money generating service (that makes more money then selling the cards).
This way there's no risk of: A competitor out competing them, not successfully developing a profitable product, "the ai bubble popping", stagnating development, etc.
There's also the advantage that this capital has allowed them to buy up most of TSMC's production capacity, which limits the competitors like Google's TPUs.
There is an inherent trade off between model size and quality. Quantization reduces model size at the expense of quality. Sometimes it's a better way to do that than reducing the number of parameters, but it's still fundamentally the same trade off. You can't make the highest quality model use the smallest amount of memory. It's information theory, not sorcery.
Quantization is essential for me since a 7B model won't fit on my RTX 2060 with only 6GB of VRAM. It allows me to compress the model so it can run on my hardware.
Soon the GPU and its associated memory will be on different cards, as once happened with CPUs. The day of the GPU with ram slots is fast approaching. We will soon plug terabytes of ram into our 4090s, then plug a half-dozen 4090s into a raspberry PI to create a Cronenberg rendering monster. Can it generate movies faster than Pixar can write them? Sure. Can it play Factorio? Heck no.
If you don't care about bandwidth you can already have a GPU access terabytes of memory across the PCIe bus, but it's too slow to be useful for basically anything. Best case you're getting 64GB/sec over PCIe 5.0 x16, when VRAM is reaching 3.3TB/sec on the highest end hardware and even mid-range consumer cards are doing >500GB/sec.
Things are headed the other way if anything, Apple and Intel are integrating RAM onto the CPU package for better performance than is possible with socketed RAM.
Thinking on the classic neural network for example, each column of nodes would only need to talk to the next column. You could group several columns per GPU and then each would process its own set of nodes. While an individual job would be slower, you could run multiple tasks in parallel, processing new inputs after each set of nodes is finished.
What's more likely is hybrid systems. Your basic desktop CPU gets e.g. 8GB of HBM, but then also has 16GB of DRAM in slots. Another CPU/APU model that fits into the same socket has 32GB of HBM (and so costs more), which you could then combine with 128GB of DRAM. Or none, by leaving the slots empty, if you want entirely HBM. A server or HEDT CPU might have 256GB of HBM and support 4TB of DRAM.
AMD still suffers from limited resources and doesn't seem willing to spend too much chasing a market that might just be a temporary hype, Google's TPUs are a pain to use and seem to have stalled out, and Intel lacks commitment, and even their products that went roughly in that direction aren't a great match for neural networks because of their philosophy of having fewer more complex cores.
https://nvidia.custhelp.com/app/answers/detail/a_id/5490/~/s...
You’ll be able to get higher resolution but slowly. Or pay the $2800 for a 5090 and get high res with good speed.
Last I saw they performed really poorly, like lower single digits t/s. Don't get me wrong they're probably a decent value for experimenting with it, but is flat out pathetic compared to an A100 or H100. And I think useless for training?
That might seem small compared to LLMs, but it isn't small in absolute terms.
GPUs need a decent virtual memory system though. The current "it runs or it crashes" situation isn't good enough.
If you size the browser window right, paging with the arrow keys (so the document doesn't scroll) you'll see (eg, pages 20-21) the textures of the parrot's feathers are almost identical to the textures of bark on the tree behind the panda bear, or the forest behind the red panda is very similar to the undersea environment.
Even if I'm misunderstanding something fundamental here about this technique, I still find this interesting!
2.1 didn't have adoption because people didn't want to deal with the open replacement for CLIP. Or possibly because everyone confused 2.0 and 2.1.
how exactly did the community deal with it? interested to learn how to unlearn safety
>>>Its in sizes from 800m to 8b parameters now, will be all sizes for all sorts of edge to giant GPU deployment.
--
Can you fragment responses such that if an edge device (mobile app) is prompted for [thing] it can pass tokens upstream on the prompt -- Torrenting responses effectively - and you could push actual GPU edge devices in certain climates... like dens cities whom are expected to be a Fton of GPU cycle consumption around the edge?
So you have tiered processing (speed is done locally, quality level 1 can take some edge gpu - and corporate shit can be handled in cloud...
----
Can you fragment and torrent a response?
If so, how is that request torn up and routed to appropriate resources?
BOFH me if this is a stupid question? (but its valid for how we are evolving to AI being intrinsic to our society so quickly.)
can someone explain how negation is currently done in stable diffusion? and why cant we do it in text LLMs?
https://pbs.twimg.com/media/GG8mm5va4AA_5PJ?format=jpg&name=...
This fascinated me when SD was first released, so I tested a whole bunch of scenarios. While it's quite easy to find situations that don't provide accurate results and produce all manner of glitches (some of which you can use to detect some SD-produced images), the results are nearly always convincing at a quick glance.
The reason it knows this is that this is how any light in a real photograph works, not just CGI.
Or if your prompt was “A green triangle looking at itself in the mirror” then early generation steps would have two green triangle like shapes. It doesn’t need to know about the concept of light reflection. It does know about composition of an image based on the word mirror though.
What if you can | a scene to a model and just have it calc all the ray-paths and then | any color/image... if you pre-calc various ray angles, you can then just map your POV and allow for the volume as it pertains to your POV be mapped with whatever overlay you want.
Here is the crazy cyberpunk part:
IT (whatever 'IT' is) keeps a lidar of everything EVERYONE senses in that space and can overlap/time/sequence anything about each experience and layer (baromoter/news/blah tied to that temporal marker)
Micro resolution of advanced lidar is used in signature creation to ensure/verify/detect fake places vs IRL.
Secret nodes are used to anti-lidar the sensors... so a place can be hidden from drones attempting to map it.
These anonolies are detectable thou, and GIS experts with terra forming skills are the new secOPs.
Fn dorks.
-- so, you already have an asset, lets say its a CUBOID room - with walls and such of wood texture_05.png
Also, "pipe" isn't considered harmful terminology (yet) just FYI. I was confused seeing the "|" mononym in it's place.
I was being lazy....
But I realize you are correctin the mirroring - I immediately thought it was ray tracing the green hue from the reflection onto a surface that could see it...
Inference is far more efficient - however - it would be really interesting to know HOW an AI 'thinks' about such reflections?
Whats the current status of AIs documenting themselves?
I imagine this doesn't look impressive to anyone unfamiliar with the scene, but this was absolutely impossible with any of the older models. Though, I still want to know if it reliabily does this--so many other things are left to chance, if I need to also hit a one-in-ten chance of the composition being right, it still might not be very useful.
We introduce Diffusion Transformers (DiTs), a simple transformer-based backbone for diffusion models that outperforms prior U-Net models and inherits the excellent scaling properties of the transformer model class. Given the promising scaling results in this paper, future work should continue to scale DiTs to larger models and token counts. DiT could also be explored as a drop-in backbone for text-to-image models like DALL E 2 and Stable Diffusion.
Afaict the answer is that combining transformers with diffusers in this way means that the models can (feasibly) operate in a much larger, more linguistically-complex space. So it’s better at spatial relationships simply because it has more computational “time” or “energy” or “attention” to focus on them.Any actual experts want to tell me if I’m close?
Generating good images is easy but generating good images with very specific instructions is not. For example, try getting midjourney to generate a shot of a road from the side (ie standing on the shoulder of a road taking a photo of the shoulder on the other side with the road crossing frame from left to right)...you'll find midjourney only wants to generate images of roads coming at the "camera" from the vanishing point. I even tried feeding an example image with the correct framing for midjourney to analyze to help inform what prompts to use, but this still did not result in the expected output. This is obviously not the only framing + subject combination that model(s) struggle with.
For people who use image generation as a tool within a larger project's workflow, this hurdle makes the tool swing back and forth from "game changing technology" to "major time sink".
If this example prompt/output is an honest demonstration of SD3's attention to specificity, especially as it pertains to framing and composition of objects + subjects, then I think its definitely impressive.
For context, I've used SD (via comfyUI), midjourney, and Dalle. All of these models + UIs have shared this issue in varying degrees.
> I often find myself having to resort to generating all of the elements I want out of an image separately and then comp them together with photoshop. This isn't a bad workflow, but it is tedious
The models should be developed to accelerate this then.
ie you should be able to say layer one is this text prompt plus this camera angle, layer two is some mountains you cheaply modeled in Blender, layer three is a sketch you drew of today's anime girl.
On the other hand, SD has just not been on the level of the quality of images I get from Midjourney. The people who counter this I don't think know what they are talking about.
Can't wait to try this.
EDIT: I think the best approach is simply to separate out the terms in separate phrases, as that gets more-or-less 100% accuracy https://imgur.com/a/JGjkicQ
That said, we should acknowledge the point of all this: SD3 is just incredibly incredibly impressive.
It has a lot of difficulty with the orientation of the cat and dog, and by the time it gets them in the right positions, the triangle is lost.
The transformer layers perform self-attention between all pairs of patches, allowing the model to build a rich understanding of the relationships between areas of an image. These relationships extend into the dimensions of the conditioning prompts, which is why you can say “put a red cube over there” and it actually is able to do that.
I suspect that the smaller model versions will do a great job of generating imagery, but may not follow the prompt as closely, but that’s just a hunch.
One of the neat things they do with the diffusion transformer is to enable creating smaller or larger models simply by changing the patch size. Smaller patches require more Gflops, but the attention is finer grained, so you would expect better output.
Another neat thing is how they apply conditioning and the time step embedding. Instead of adding these in a special way, they simply inject them as tokens, no different from the image patch tokens. The transformer model builds its own notion of what these things mean.
This implies that you could inject tokens representing anything you want. With the U-Net architecture in stable diffusion, for instance, we have to hook onto the side of the model to control it in various sort of hacky ways. With DiT, you would just add your control tokens and fine tune the model. That’s extremely powerful and flexible and I look forward to a whole lot more innovation happening simply because training in new concepts will be so straightforward.
Before: Evaluate the image in a little region around each pixel against the prompt as a whole -- e.g. how well does a little 10x10 chunk of pixels map to a prompt about a "red sphere and blue cube". This is problematic because maybe all the pixels are red but you can't "see" whether it's the sphere or the cube.
After: Evaluate the image as a whole against chunks of the prompt. So now we're looking at a room, and then we patch in (layer?) a "red sphere" and then do it again with a "blue cube".
Is that roughly the idea?
Tried all morning and ChatGPT could not do it.
Or is the highway literally being held by a humanoid plane?
Likewise with Mistral, you don't get half a billion in funding and a two billion valuation on the assumption that you'll keep giving the product away for free forever.
I use Midjourney a lot and (based on the images in the article) it’s leaps and bounds beyond SD. Not sure why I would switch if they are both locked down.
I think the ability for people to adopt models made SD more successful than any other model for image synthesis in the first place.
Similarly how consumer PCs drove innovation towards faster hardware.
I believe it to be the reference of image synthesis for that matter, so "better" is a bit blurry.
Although I don't understand the criticism of the images in question. Without a prompt comparison, it is impossible to compare image synthesis. What are examples of images that are beyond these?
I am using Midjourney to basically create images in particular artistic styles (e.g., “painting of coffee cup in ukiyo-e style”) and that works very well. I am interested in SD for creating images based on artwork that isn’t indexed by Midjourney, though, as some of the more obscure artists aren’t available.
Of course such sites are heavily biased towards content that is popular, but you will also find quite specific models if you search for certain styles.
I don’t mind that they don’t want to let you generate nsfw images but their detector is hopelessly broken, it once censored a cube, yes a cube...
Here is the red cube it censored because my innocent eyes wouldn't be able to handle it; https://archerx.com/censoredcube.png
What they are achieving with the over zealous safety issues are driving developers to on demand GPU hosts that will let them host their own models, which also opens up a lot more freedom. I wanted to use the stability AI api as my main source for Stable Diffusion but they make it really really hard especially if you want use it as part of your business.
I would much rather have this than a company releasing models this size into the wild without any safety checks whatsoever.
Until then, we must view this “safety” as both a scapegoat and a vector for social engineering.
Is it unlikely? Sure, but worth validating.
I posed the worst-case scenario of generating actual CSAM in response to your question, "What particular image that you think a random human will ask the AI to generate, which then leads to concrete harm in the real world?"
https://www.404media.co/laion-datasets-removed-stanford-csam...
Besides, and this is a serious question, what's the harm of a model accidentally generating CSAM? If you weren't intending to generate these images then you would just discard the output, no harm done.
Nobody is forcing you to use a model that might accidentally offend you with its output. You can try "aligning" it, but you'll just end up with Google Gemini style "Sorry I can't generate pictures of white people".
And yeah I think we should care, for a lot of reasons, but a big one is just trying to stay well within the law.
[0] https://www.404media.co/laion-datasets-removed-stanford-csam...
What I had in mind were regular general purpose models which I've played around with quite extensively.
Add to that, parents who want to avoid having their kids generate sexual content would now need to prevent their kids from using this tool because it can create it randomly, limiting SD usage to kids 18+ (which is probably something else Stability AI does not want to deal with.)
It's definitely a balance between going overboard and having restrictions though. I haven't used SD in several months now so I'm not sure where that balance is right now.
To whom? SD's reputation, perhaps - but that ship has already sailed with 1.x. That aside, why is generated porn threatening? If anything, anti-porn crusaders ought to rejoice, given that it doesn't involve actual humans performing all those acts.
You can have your own opinion on it, but surely you can see the issue here?
I've noticed that SDXL does something a little odd. For a given prompt it essentially decides what race the subject should be without the prompt having specified one. You generate 20 images with 20 different seeds but the same prompt and they're typically all the same race. In some cases they even appear to be the same "person" even though I doubt it's a real person (at least not anyone I could recognize as a known public figure any of the times it did this). I'm kind of curious what they changed from SD 1.5, which didn't do this.
And safe doesn't mean "lower than 1/10^6 chance of ending humanity", safe means shoddily implemented curtailing to idpol + fundamentalist level moral aversion towards human sexuality
There is such a great liberal fear of being perceived as any of the negative -ists and -isms that the pendulum swings to the other extreme where the left horseshoe toe meets its rightmost brother, which is why SD and Google's new toy rewrite ancient European history to include POC's and queer people.
I wonder if they are afraid of the same debacle as google AI and what they mean by "safety" is actually heavy bias against white people and their culture like what happened with Gemini.
The lie originates with a Communist race hustler named Noel Ignatiev, also known for publishing Race Traitor magazine. A thoroughly unpleasant person.
As opposed to truthful liars?
If this next version is just as bad, I'm going to stop using Stability APIs. Are there any other text-to-image services that offer similar value and quality to Stable Diffusion without the overzealous blurring?
Edit:
Example prompt's like "Matte portrait of Yennefer" return 8/9 blurred images [1]
I tend to lean towards SD1.5 for this reason—I'd rather put in the effort to get a good result out of the lesser model than fight with a black box censorship algorithm.
EDIT: See the replies below. I might just have been holding it wrong.
EDIT: based on the other reply, I think I understand what you're suggesting, and I'll definitely take a look next time I run it.
Same goes for upscalers, of course.
If you are using invoke, try XL.
If you want to really dial into a specific style or apply a specific LORA, use 1.5.
Is the OSS'd version of SDXL less restrictive than their API hosted version?
After SD1.5 they started directly modifying the dataset.
it's only other users who "restore" the porno.
and that's what we're discussing. there's a real concern about it as a public offering.
they both don't want to offer anything that's legally dubious and it's not hard to understand why.
The models being open sourced makes them very easy to turn into the most deprived porno machines ever conceived. And they are.
It is in no way a meaningful barrier to what people can do. That’s the benefit of open source software.
Gemini demonstrated a product I do not want to use and I am aware about the requirements of corporate contexts, although I think the safety mechanisms should be in the hand of users.
Google optimized for advertisers, but I am not interested in such content as it provides little value.
No large scale model maker is going to put out public models for B2B with dubious use cases.
what the problem is: OpenAI, facebook, Google are not curating the data sets. you're arguing they shouldn't put controls after the fact. but what you actually want is them to use quality datasets.
[1] Google: https://www.google.com/search?sca_esv=a930a3196aed2650&q=yen...
[2] Bing via Ecosia: https://www.ecosia.org/images?q=yennefer%20witcher%203%20gam...
[3] Bing: https://www.bing.com/images/search?q=yennefer+witcher+3+game...
[4] DDG: https://duckduckgo.com/?va=e&t=hj&q=yennefer+witcher+3+game&...
[5] Yippy: https://www.alltheinternet.com/?q=yennefer+witcher+3+game&ar...
[6] Dogpile: https://www.dogpile.com/serp?qc=images&q=yennefer+witcher+3+...
I’ve had some fair share of frustation with DallE as well when trying to generate weapon images for game assets. Had to tweak a lot of my prompt
The fact that they have censorship values is scary. But the fact that those are different is better than the alternative.
What exactly does this mean? Will we be able to see all of the "safeguards" and access all of the technology's power without someone else's restrictions on them?
Think of childern! We must stop people from generating porn!
Kind of a dud for an announcement.
will the model also be able to produce good photographs, technical drawings, and other graphical media?
It has the weird art style because that's what looks the most "aesthetic". And because it doesn't actually have nearly as good enough data as you'd think it does.
Sora looks like it could be better.
I'd want a model that can draw website designs and other UIs well. So I give it a list of things in the UI, and I get back a bunch of UI design examples with those elements.
"We believe in safe, responsible knife practices. This means we have taken and continue to take reasonable steps to prevent the misuse of Big Knife by bad actors."
I'm only now investigating using AI to increase velocity in my projects, and the field is moving so fast, i'm a bit outdated.
From the FAQ: "v0 is a generative user interface system by Vercel powered by AI. It generates copy-and-paste friendly React code based on shadcn/ui and Tailwind CSS that people can use in their projects"
(The second is what we did for https://gwern.net/dropcap because the PNG->SVG filesizes & quality were just barely acceptable for our web pages.)
midjourney 6 can be completely photorealistic and include valid text, but also sometimes adds bad text. it's not much, but having to use an image editor for that is still annoying. for creating marketing material, getting perfect text every time and never getting bad text would be amazing
https://www.youtube.com/watch?v=_7rMfsA24Ls https://course.fast.ai/Lessons/part2.html
His whole blog is fantastic. If you want more background (e.g. how transformers work) he's got all the posts you need
The older ones have drawbacks like not being able to spell.
Which goes far more towards the idea that safety isn’t a desirable feature to a lot of AI users.
I mean, SDXL is great. Until you’ve had a chance to actually use this model, isn’t calling it out for some imagined offence that may or may not exist seems like you’re drinking some Kool-aid rather than responding to something based in concrete actual reality.
You get access to it… and it does the google thing and puts people of colour in every frame? Sure, complain away.
You get access to it, you can’t even generate pictures of girls? Sure. Burn the house down.
…you haven’t even seen it and you’re already bitching about it?
Come on… give them a chance. Judge what it is when you see it not what you imagine it is before you’ve even had a chance to try it out…
Lots of models, free, multiple sizes, hot damn. This is cool stuff. Be a bit grateful for the work they’re doing.
…and even if sucks, it’s open. If it’s not what you want, you can retune it.
It’s been 6 months and it still isn’t there. SD3 is going to be quite awhile if they’re baking “safety” in even harder.
I wish I had something more clever to comment on it. I know what they’re doing which is cool and why which is, IDK, live and let live and enjoy your own kink. It just a little funny some of the most work put into in the fine tuning models.. is from the pony community.
So all I have is…
:/
Sadly the elite ai hackers have shuffled off into private discords.
We are just trying to satisfy your values though ponies and friendship.
I used rundiffusion to play around with a bunch of different open source software quickly and easily with pre-downloaded models after getting annoyed at my laptop GPU. But once I settled on one particular implementation and started spending a lot of time in it, it no longer made sense to repeatedly pay every hour for an initial ease-of-setup.
The only real ongoing benefit was rundiffusion came with a bunch of models pre-downloaded so swapping between them was quick. But you can use UI addons like the CivitAI browser to download models automatically through automatic1111, and you'll likely want to go beyond what they predownload to the instance for you anyway.
The downside to running on the cloud directly is having to manage the running/stopped state of the instance yourself. I haven't ever left it running when I was done with an instance, but I could see that as a risk. CLI commands and scripting can make that faster than logging into a website which does it for you automatically, but it's extra effort.
I thought about building an AMI and putting it up on AWS marketplace, but it looks like there are a few options for that already. I don't know how good they are out of the box, as I haven't used them. But if spending 20 minutes once to get software running on a Linux instance is truly the only barrier to reducing cost, those prebuilt AMIs are a decent intermediary step. They're about $0.10/hour on top of server costs. I skipped straight to installing the software myself, but even an extra $0.10/hour overhead would be better than paying double..
1. Leaving instances running when they're not being used
and
2. Deviation from default behavior that results in accumulation of storage volumes you don't want or need (low likelihood but something to watch for initially).
For 1:
If you leave the instance running you'll keep getting charged the hourly rate. Not really unexpected, but you have to notice it yourself or set an alarm.
There are a few tricks to reduce likelihood of this happening and to limit charges if it does happen anyway:
a. Prevention: Make your own little auto-stop script for the instance like rundiffusion has. Maybe make it into the launch sequence too, so you run a script, it launches the instance, then starts a timer. If the timer counts down all the way with you jiggling it, it stops the instance.
b. Mitigation: Create an alarm on the instance with the action to 'Stop' the instance when the alarm is triggered. Set the trigger for the alarm to be something like 'Max CPU usage has been less than 4% for a consecutive hour'.
c. Mitigation: Use AWS' Instance scheduler to automatically stop the instance
d. Mitigation: Billing budgets with associated action to stop instances -- kind of like the alarms but triggered based on costs
For 2:
It's probably a non-issue. You'll likely not have a problem because you'll start and stop the same instance most of the time instead of creating and deleting new instances. In which case, gp3 SSD storage is $0.08/gb over a month, charges on a 200gb storage volume you keep around all the time and use is only like $16 for the month. There are benefits so it's likely worthwhile.
BUT, be careful if you create and terminate lots of instances instead of stopping and starting the same instance. There's a small possibility of accumulating extra storage volumes you don't need, without realizing it.
By DEFAULT that's not a problem. AWS will delete the attached storage volumes for an instance after you Terminate (not stop) the instance. The problem comes from changing the default behavior.
That change can be configured in the AMI you use to launch an instance (by whoever created the AMI), or by you when launching an instance. The storage volumes are not deleted automatically then you have to do it manually to stop them from continuing to generate charges. Which you might not notice right away..
Keep that in mind, but that scenario requires a fairly unlikely chain of requirements to get to the point of bill bloat:
- If you use a pre-built AMI from marketplace (such as one with stablediffusion preinstalled) which is configured to KEEP storage volumes upon termination instead of using AWS default of deleting them,
- and if you do a 'Create' and 'Terminate' instead of 'Start/Stop' so you're using a LOT of instances instead of a few,
- and if you don't notice the setting during launch,
- and if you don't see all the extra volumes sitting around...
[0]: https://twitter.com/EMostaque/status/1760660709308846135
Wouldn't this v3 supersede the StableCascade work?
Did they announce it because a team had been working on it and they wanted to push it out to not just lose it as an internal project, or are there architectural differences that make both worthwile?
It’s a really smart architecture and I think is fertile ground for stacking on new things like DiT.
SD3 seems to be more towards SOTA, not sure why Cascade took so long to get out, seemed to be up and running months ago
Nevertheless, it’s frustrating that the industry is fragmenting to a variety of licenses where you have to read the fine print and often licensing information isn’t announced until final release.
Also, I was blown away by the "Stable Diffusion" written on the side of the bus.
Not that weird.
Which takes all the behind the scenes steps, not just the technical ones.
There was a country that had the best mathematicians, the best physicists, the best metallurgists in the world. But that country was very poor. It’s called the Soviet Union. But when you took one of these mathematicians or physicists, who was smuggled out or escaped, put him on a plane and brought him to Palo Alto. Within two weeks, they were producing added value that could produce great wealth.
What comes first is markets. If you have great technology without markets, without a market-friendly economy, you’ll get nowhere. But if you have a market-friendly economy, sooner or later the market forces will give you the technology you want.
And that my friend, simply won't come from an office paralyzed by internal politics of fear and conformity. Don't get it twisted.
People always say Google is "behind", I don't believe they're behind in a capabilities sense, which IMO is what the parent is implying. They've decided to make a PC product, which I wouldn't say is inferior to anything else if you're the kind of person who is into PC culture.
There might be some difficult internal politics to work through, but there is no way Google is hamstrung forever by this.
Go calm my friend.
I'm not.
> There might be some difficult internal politics to work through, but there is no way Google is hamstrung forever by this.
The technical prowess is irrelevant. Whichever of the companies in the AI race excises their PC demons will actually ship useful things and break ahead. The talent will follow the market.
Google might very well be hamstrung forever by this and other internal politics. This dynamic has played out over and over again.
See also: IBM, Xerox, HP, Nokia. 'etc.
Like as if there would be many companies in the world who have the resources, know how and expertise to pull off event the PC version of the product.
Stability does have an LLM, but it's not provided in a unified framework like Gemini is.
I wonder how far ahead the internal versions are?
Would investors stop giving them money? Would users sue that they now had PTSD after looking at all the 'unsafe' outputs? Would regulators step in and make laws banning this 'unsafe' AI?
What is it specifically that company management is worried about?
One example of that would be if your model was being used to spot criminals in video footage and it turns out that the bias of the model picks one socioeconomic group over another. Most western nations have laws protecting the public against that kind of abuse (albeit they're not applied fairly) and the fines are pretty steep.
I'm glad tech orgs are for once thinking about what they're building before putting out society-warping democracy-corroding technology instead of move fast break things.
I'd be on your side if any of them actually chose to keep their technology in the lab instead of tossing it out into the world and gobbling up investment dollars as fast as they could.
Software that promotes the unchecked spread of propaganda, conspiracy theories, hostility, division, institutional mistrust and so on: A-OK.
Software that might show a boob: Totally irresponsible and deserving of harsh regulation.
Also, where do we draw the line? Should Photoshop stop you from manipulating human body because it could be used for porn? Why stop there, should text editors stop you from writing about sex or describing human body because it could be used for "abuse". Should your comment be removed because it make me imagine Taylor Swift without clothes for a brief moment?
(This applies to all AI discussions)
I don't know much about the initial scandal, but I was under the impression that there was only a small number of those images, yet that didn't change the situation. I just fail to see how quantity factors into anything here.
Because you can overload any online discussion / sphere with that. There were so many that X effectively banned searching for her at all because if you did, you where overwhelmed by very extreme fake porn. Everybody can do it with very low entry barrier, it looks very believable, and it can be generated in high quantities.
We shouldn't have clamped down on photoshop, but realisticly two things would be nice in your theoretical case, usage restrictions and public information building. There was no clear cut point where photoshop was so mighty you couldn't trust any picture online. There were skills to be learned and people could identify the trickery, and it was on a very small scale and gradual. And the photo trickery was around for ages, even Stalin did it.
But creating photorealistic fakes in an automated fashion is completely new.
And in fairness to generative AI, even nowadays it feels like getting to a point of true photorealism takes some effort, especially if the goal is letting it just run nonstop with no further curation. And getting a local image generator to run at all on your computer (and having the hardware for it) is also a bar that plenty of people can't clear yet. Photoshop is kind of different in that making more believable things requires a lot more time, effort and knowledge - but the idea that any image online can be faked has already been ingrained in the public consciousness for a very long time.
Think of it this way, if one out of every ten phone calls you get is spam, you still have a pretty useable phone. Three orders of magnitude different and 1 out of every 100 calls is real and the system totally breaks down.
Generative AI makes generating realistic looking fakes ~1000x easier, its the one thing its best at.
but that's not dangerous. It's definitely worthy of unlocking the cages of the attack lawyers but it's not dangerous. The word "safety" is being used by big tech to trigger and gas light society.
The safety discussion is proceeding very much like it did for movies, music, and video games.
https://www.whitehouse.gov/briefing-room/presidential-action...
Not engaging in this will indeed lead to bad laws, sanctions and more as well as not fulfilling our societal obligations of ensuring this amazing technology is used for as positive outcomes as possible.
Stability AI was set up to build benchmark open models of all types in a proper way, this is why for example we are one of the only companies to offer opt out of datasets (stable cascade and SD3 are opted out), have given millions of supercompute hours in grants to safety related research and more.
Smaller players with less uptake and scrutiny don't need to worry so much about some of these complex issues, it is quite a lot to keep on top of, doing our best.
Can you define what you mean by "societal and other considerations"? If not, why not?
You sound like many authoritarian regimes.
As with all hype techs, even the most talented management are barely literate in the product. When talking about their new trillion $ product they must take their talking points from the established literature and "fake it till they make it".
If the other big players say "billions of parameters" you chuck in as many as you can. If the buzz words are "tokens" you say we have lots of tokens. If the buzz words are "safety" you say we are super safe. You say them all and hope against hope that nobody asks a simple question you are not equipped to answer that will show you dont actually know what you are talking about.
The rest of the world is also like that. You can make a thing that hurts your existing business. Spinning off the brand is probably Google's best bet.
I don't think those companies being cautious is necessarily a bad thing even for AI enthusiasts. Open source models will quickly catch up without any censorship while most of those public attacks are concentrated into those high profile companies, which have established some defenses. That would be a much cheaper price than living with some unreasonable degree of regulations over decades, driven by populist politicians.
They're probably more concerned about generated images of politicians in 'interesting' sitations going viral than they are about porn/gore etc.
The old one can't.
"a woman chasing a bear, pursuit"
The artistic value is something you have to add with a good prompt with artistic vision. These images are probably the AI equivalent of "programmer art". It fulfills its function, but lacks aesthetic considerations. I wouldn't attribute that to the model just yet.
Perhaps eventually, once every forum has been assigned a trust-and-safety team and word processor has been aligned and most normal people have no need for communication outside the Metaverse (TM) in their daily lives, we will also come around to reviewing the necessity of teaching kids to write, considering the epidemic of hateful graffiti and children being caught with handwritten sexualised depictions of their classmates.
What makes you think those who’ve worked hard over a lifetime to provide (with no compensation) the vast amounts of data required for these — inferior by every metric other than quantity — stochastic approximations of human thought should feel empowered?
I think the genAI / printing press analogy is wearing rather thin now.
He did not get into the profession to make money. He did it out of passion and died poor. Artists are not being tricked by the promise of wealth. You will get a cloned style if you can't afford the real artist making it and if the commission goes to a computer how is that not the same as plagerism by a human? Artists were not being paid well before. The anime industry has proven the endpoint of what happens to artists as a profession despite their skills. Chess still exists despite better play by machines. Art as a commercial medium has always been tainted by outside influences such as government, religion and pedophilia.
In the end, drawing wasn't going to survive in the age of vector art and computers. They are mainly forgettable jpgs you scroll past in a vast array like DeviantArt.
Anyway, my point isn’t that ‘AI is evil and must be stopped’; it’s that it doesn’t feel ‘intellectually empowering’. I (in my personal work) can’t get anything done with ChatGPT that I can’t on my own, and with less frustration. We’ve created machines that can superficially mimic real work, and the world is going bonkers over it. The only magic power these systems have is sheer speed: they can output reams and reams of twaddle in the time it takes me to make a cup of tea. And no doubt those in bullshit jobs are soon going to find out.
My argument might not be what you expect from someone who is sad to see the way artists’ lives are going: if your work is truly capable of being replaced by a large language model or a diffusion model, maybe it wasn’t very original to begin with.
The sad thing is, artists who create genuinely superior work will still lose out because those financially enabling them will think (wrongly) that they can be replaced. And we’ll all be worse off.
You keep going back to value and finances. The less money is in it the better. Art isn't good because it's valuable, unless you were only interested in it commercially.
Of course not; I’m certainly not suggesting so. But I do think money is important because it is what has enabled artists to do what they do. Without any prospect of monetising one’s art, most of us (and I’m not an artist) would be out working in the potato fields, with very little time to develop skills.
Commercial movies have lots of CG, big budgets and famous actors while small budget indie movies have been exploding despire their weaker technical specialities. Noah's ark was made by amateurs while the titanic was made by experts.
And the metric of "beating most of our existing metrics so we had to rewrite the metrics to keep feeling special, but don't worry we can justify this rewriting by pointing at Goodhart's law".
The only reason the question of compensating people for their input into these models even matters is specifically because the models are, in actual fact, good. The bad models don't replace anyone.
This is needlessly provocative, and also wrong. My metrics have been the same from the very beginning (i.e. ‘can it even come close to doing my work for me?’). This question may yet come to evaluate to ‘yes’, but I think you seriously underestimate the real power of these models.
> The only reason the question of compensating people for their input into these models even matters is specifically because the models are, in actual fact, good.
No. They don’t need to be good, they simply need to fool people into thinking they’re good.
And before you reflexively rebut with ‘what’s the difference?’, let me ask you this: is the quality of a piece of work or the importance of a job and all of its indirect effects always immediately apparent? Is it possible for managers to short term cost-cut at the expense of the long term? Is it conceivable that we could at some point slip into a world in which there is no funding for genuinely interesting media anymore because 90% of the population can’t distinguish it? The real danger of genAI is that it convinces non-experts that the experts are replaceable when the reality is utterly different. In some cases this will lead to serious blowups and the real experts will be called back in, but in more ambiguous cases we’ll just quietly lose something of real value.
Perhaps; this is something I find annoying enough that my responses may be unnecessarily sharp…
> and also wrong. My metrics have been the same from the very beginning (i.e. ‘can it even come close to doing my work for me?’). This question may yet come to evaluate to ‘yes’, but I think you seriously underestimate the real power of these models.
Okay then. (1) your definition is equivalent to "permanent mass unemployment" because if it can do your work for you, it can also do your work for someone else, (2) you mean either "over-estimate" or "real limits of these models", and the only reason I even bring up what's obviously a minor editing issue that I fall foul of myself on many comments is that this is the kind of mistake that people pick up on as evidence of the limits of AI — treating small inversions like this as evidence of uselessness.
> Is it conceivable that we could at some point slip into a world in which there is no funding for genuinely interesting media anymore because 90% of the population can’t distinguish it?
As written, what you describe is tautologically impossible. However, assuming you mean something more like "genuinely novel" rather than "interesting", absolutely! 100% yes. There's also loads of ways this could permanently end all human flourishing (even when used as a mere tool e.g. by dictators for propaganda), and some plausible ways it can permanently end all human existence (it's a safe bet someone will ask it to and try to empower it to this end, the question is how far they get with this).
> The real danger of genAI is that it convinces non-experts that the experts are replaceable when the reality is utterly different.
Despite the fact that the best models ace tests in medicine and law, the international mathematical olympiad, leetcode, etc., the fact there are no real tests for how good someone is after a few years of employment means both your point and mine can be true simultaneously. I'm thinking the real threat current LLMs pose to newspapers is that they fully automate the Gell-Mann Amnesia effect, even though they beat humans on every measure I had of intelligence when I was growing up, and depending on which measure exactly either all of humanity together by many orders of magnitude, or at worst putting them somewhere near the level of "rather good student taking the same test".
> In some cases this will lead to serious blowups and the real experts will be called back in, but in more ambiguous cases we’ll just quietly lose something of real value.
Hard disagree about "quiet loss". To the extent that value can be quantified, even if only by surveying humans, models can learn it. Indeed, this is already baked into the way ChatGPT asks you for feedback about the quality of the answers it generates. To the extent we lose things, it will be a very loud and noisy loss, possibly literally in the form of a nuke going off.
This wouldn't happen because employment effects are mainly determined by comparative advantage, i.e. the resources that could be used to "do your job" will instead be used to do something they're more suited to.
(Not "that they're better at". it's "more suited to". You do not have your job because you're the best at it.)
If I imagine a world where every task that any human can perform can also be done at world expert level — let alone at a superhuman level — by a computer/robot (with my implicit assumption "cheaply"), I can't imagine why I would ever choose the human option. If the comparative advantage argument is "the computer/robot combination will always be priced at exactly the level where it's cost-competitive with a human, in order that it can extract maximum profit", I ask why there won't be many AI/robots competing with each other for ever-smaller profit margins?
[0] AI and robotics are not the same things, one is body the other mind, but there's a lot of overlap with AI being used to drive robots, LLMs making it easier to define rewards and for the robots to plan; and AI also get better by having embodiment (even if virtual) giving them real world feedback.
[1] https://www.wolframalpha.com/input?i=5+months+*+log2%28mass+...
Lot of hidden assumptions here. How does "operating at human level" (an assumption itself) imply the ability to do this? Humans can't do this.
We very specifically can't do this, we have sexual reproduction for a good reason.
(Also, since your scenario also has the robots working for free, they would instantly run out of resources to reproduce because they don't have any money. Similarly, an AGI will be unable to grow exponentially and take over the world because it would have to pay its AWS bill.)
> If I imagine a world where every task that any human can perform can also be done at world expert level — let alone at a superhuman level — by a computer/robot (with my implicit assumption "cheaply"), I can't imagine why I would ever choose the human option.
If the robot performs at human level, and it knows you'll always hire it over a human, why would it work for cheaper?
If you can program it to work for free, then it's subhuman.
If you're imagining something that's superhuman in only ways that are bad for you and subhuman in ways that would be good for you, just stop imagining it and you're good.
Operating at human level is directly equivalent to "can it even come close to doing my work for me" when the latter is generalised over all humans, which is the statement I was criticising on the grounds of the impact it has.
> Humans can't do this.
> We very specifically can't do this, we have sexual reproduction for a good reason.
Tautologically, humans operate at human level.
If you were responding to «"make a better version of itself" until that process hits a limit» — we've been doing, and continue to do, that with things like "education" and "medicine" and "sanitation". We've not hit our limits yet, as we definitely don't fully understand how DNA influences intelligence, nor how to safely modify it (plenty of unsafe ways to do so, though).
If you were responding to «followed by "maximise how many of you exist" until it runs out of resources», that's something all living things do by default. Despite the reduced fertility rates, our population is still rising.
And I have no idea what your point is about sexual reproduction, because it's trivial to implement a genetic algorithm in software, and we already do as a form of AI.
> (Also, since your scenario also has the robots working for free, they would instantly run out of resources to reproduce because they don't have any money. Similarly, an AGI will be unable to grow exponentially and take over the world because it would have to pay its AWS bill.)
First, I didn't say "for free", I was saying "competing with each other such that the profit margin tends towards zero", which is different.
Second, money is an abstraction to enable cooperation, it is not the resource itself. Money doesn't grow on trees, but apples do: just as plants don't use money but instead takes minerals out of the soil, carbon out of the air, and water out of both, so too a robot which mines and processes some trace elements, silicon, and iron ore into PV and steel has those products as resources even if it doesn't then go on to sell them to anyone. Inventing the first VN machine involves money, but only because the humans used to invent all the parts of that tech themselves want money while working on the process.
AI may still use money to coordinate, because it's a really good abstraction, but I wouldn't want to bet against superior coordination mechanisms replacing it at any arbitrary point in the future, neither for AI nor for humans.
> If the robot performs at human level, and it knows you'll always hire it over a human, why would it work for cheaper?
(1) competition with all the other robots who are trying to bid lower to get the business, i.e. Nash equilibrium of a free market
(2) I dispute the claim that "If you can program it to work for free, then it's subhuman." because all you have to do is give it a reward function that makes it want to make humans happy, and there are humans who value the idea of service as a reward all in its own right. Further, I think you are mixing categories by calling it "subhuman", as it sounds like an argument based on the value of its inner experience, where the economic result only requires the productive outputs — so for example, I would be surprised if it turned out Stable Diffusion models experienced qualia (making them "subhuman" in the moral value sense), but they're still capable of far better artistic output than most humans, to the extent that many artists are giving up on their profession (making them superhuman in the economic sense).
(3) One thing humans can do is program robots, which we're already doing, so if an AI were good enough to reach the standard I was objecting to, "can it even come close to doing my work for me" fully generalised over all humans, then the AI can program "subhuman" labour bots just as easily as we can, regardless of whether or not there turns out to be some requirement for qualia to enable performance in specific areas.
I think you have a conceptual confusion here. "Medicine" doesn't exist as an entity, and if it does, it doesn't do anything. People discover new things in the field of medicine. Those people are not medicine. (If they're claiming to be, they aren't, because of the principal-agent problem.)
> And I have no idea what your point is about sexual reproduction, because it's trivial to implement a genetic algorithm in software, and we already do as a form of AI.
Conceptual confusion again. Just because you call different things AI doesn't mean those things have anything in common or their properties can be combined with each other.
And the point is that sexual reproduction does not "make a better version of you". It forces you to cooperate with another person who has different interests than you.
Similarly, your ideas about robots building other little smaller robots who'll cooperate with each other… why are they going to cooperate with each other against you again? They don't have the same interests as each other because they're different beings.
> AI may still use money to coordinate, because it's a really good abstraction, but I wouldn't want to bet against superior coordination mechanisms replacing it at any arbitrary point in the future, neither for AI nor for humans.
Highly doubtful there could be one that wouldn't fall under the definition of money. The reason it exists is called the economic calculation problem (or the socialist calculation problem if you like); no amount of AI can be smart enough to make central planning work.
> (2) I dispute the claim that "If you can program it to work for free, then it's subhuman." because all you have to do is give it a reward function that makes it want to make humans happy
If it has a reward function it's subhuman. Humans don't have reward functions, which makes us infinitely adaptable, which means we always have comparative advantage over a robot.
> and there are humans who value the idea of service as a reward all in its own right.
It's recommended to still pay those people. That's because if you deliberately undercharge for your work, you'll run out of money eventually and die. (This is the actual meaning of efficient markets hypothesis / "people are rational" theory. It's not that people are magically rational. The irrational ones just go broke.)
Actually, it's also the reason economics is called "the dismal science". Slaveholders called it that because economists said it's inefficient to own slaves. It'd be inefficient to employ AI slaves too.
This is a strange question since augmentation can be objectively measured even as its utility is contextual. With MidJourney I do not feel augmented because while it makes pretty images, it does not make precisely the pretty images I want. I find this useless, but for the odd person who is satisfied only with looking at pretty pictures, it might be enough. Their ability to produce pretty pictures to satisfaction is thus augmented.
With GPT4 and Copilot, I am augmented in a speed instead of capabilities sense. The set of problems I can solve is not meaningfully enhanced, but my ability to close knowledge gaps is. While LLMs are limited in their global ability to help design, architect or structure the approach to a novel problem or its breakdown, they can tell local tricks and implementation approaches I do not know but can verify as correct. And even when wrong, I can often work out how to fix their approach (this is still a speed up since I likely would not have arrived at this solution concept on my own). This is a significant augmentation even if not to the level I'd like.
The reason capabilities are not much enhanced is to get the most out of LLMs, you need to be able to verify solutions due to their unreliability. If a solution contains concepts you do not know, the effort to gain the knowledge required to verify the approach (which the LLM itself can help with) needs to be manageable in reasonable time.
I am not a programmer, so none of this applies to me. I can only speak for myself, and I’m not claiming that no one can feel empowered by these tools - in fact it seems obvious that they can.
I think programmers tend to assume that all other technical jobs can be attacked in the same way, which is not necessarily true. Writing code seems to be an ideal use case for LLMs, especially given the volume of data available on the open web.
Writing it as "feel empowered" made it come across as if you meant the empowerment was illusory. My argument was that it is not merely a feeling but a real measurable difference.
It seems to me that the standard use of "empowering" implies in particular that you get more power for less effort - which in many cases tends to be democratizing, as hard-earned power tends to be accrued by a handful of people who dedicate most of their lives to pursuit of power in one form or another. With public schooling and printing, a lot of average people were empowered at the expense of nobles and clerics, who put in a lifetime of effort for the power literacy conveys in a world without widespread literacy. With AI, likewise, average people will be empowered at the expense of those who dedicated their life to learn to (draw, write good copy, program) - this looks bad because we hold those people in high esteem in a world where their talents are rare, but consider that following that appearance is analogously fallacious to loathing democratization of writing because of how noble the nobles and monks looked relative to the illiterate masses.
When writing and printing emerged, they too depended on supply chains (for paper, iron, machining) and in the case of printing capital that were far out of the reach of the individual. Their utility and overlap with other mass markets resulted in those being commoditized in short order.
Think about it: the saturation of content on the Internet has become so bad that people are having a hard time knowing what's true or not, to the point that we're having again outbreaks of preventable diseases such as measles because people can't identify what's real scientific information and what's not. Imagine what will happen when anyone can create an image of whatever they want that looks just like any other picture, or worse, video. We are not at all equipped to deal with that. We are risking a lot just for the ability to spend massive amounts of compute power on generating images. It's not curing cancer, not solving world hunger, not making space travel free, no: it's generating images.
Printing Press -> Reformation -> Thirty Years' War -> Millions Dead
I'm sure that there were lots of different opinions at the time about what kind of harm was introduced by the printing press and what to do about it, and attempts to control information by the Catholic church etc.
The current fad for 'safe' 'AI' is corporate and naive. But there's no simple way to navigate a revolutionary change in the way information is accessed / communicated.
The lesson isn't. printing press bad, it's extremist irrational belief in any entity is bad (whether it's religion, Trump, etc.).
I don't see GP blaming the printing press for that, they're merely pointing out that one enabled the other, which is absolutely true. I'm damn near a free speech absolutist, and I think the heavy "safety" push by AI is well-meaning but will have unintended consequences that cause more harm than they are meant to prevent, but it seems obvious to me that they can be used much the same as printing presses were by the extremists.
> The lesson isn't. printing press bad, it's extremist irrational belief in any entity is bad (whether it's religion, Trump, etc.).
Could not agree more
TPYOS!?
What are you a DAFT???
er.. I mean DRAFT!!!
- it was just a thought, sir..,
Would millions have died if the old religion gave way to the new one without a fight? The problem for the Vatican was that their rhetoric wasn't at top form after mentally stagnating for a few centuries since arguing with Roman pagans, so war was the only possibility to win.
(Don't forget Luther's post hoc justification of killing 100k+ peasants, but he won because he had better rhetorical skills AND the backing of aristocrats and armies. https://en.wikipedia.org/wiki/Against_the_Murderous,_Thievin... and https://en.wikipedia.org/wiki/German_Peasants%27_War)
I almost think it's the eras between that are more notable.
https://www.betterworldbooks.com/product/detail/the-coddling...
https://www.audible.com/pd/The-Coddling-of-the-American-Mind...
https://en.m.wikipedia.org/wiki/Licensing_of_the_Press_Act_1...
(For clarity, I'm joking, and I know you're also not implying any such thing. I appreciate your comment/link)
The out group used to be atheists, or gays, or witches, or republicans (in the British sense of the word), or people who want to drink. And each of Catholics and Protestants made the other unwelcome across Europe for a century or two. When I was a kid, it was anyone who wanted to smoke weed, or (because UK) any normalised depiction of gay male relationships as being at all equivalent to heterosexual ones[0]. I met someone who was embarrassed to admit they named their son "Hussein"[1], and absolutely any attempt to suggest that ecstasy was anything other than evil. I know at least one trans person who started out of the closet, but was very eager to go into the closet.
[0] "promote the teaching in any maintained school of the acceptability of homosexuality as a pretended family relationship" - https://en.wikipedia.org/wiki/Section_28
If everyone uses Hosting Service F, then at some point people will blur the lines and expect "Hosting Service F" to remove vulgar or offensive content. The lines themselves will be a zeitgeist of sorts with inevitable decisions that are acceptable to some but not all.
Can you even blame them? There are lots of ways for this to go wrong and noone wants to be on the wrong side of a PR blast.
So heavy guardrails are effectively inevitable.
edit: yes it is sarcasm, though I fear somebody will think it is in fact the right way to go.
Maybe we should make it illegal to draw or write anything without submitting it to the state for "safety" analysis.
Hopefully you are going to be absolutely shocked by the prospect of the above sentence. But as you can see, surveillance is a slippery slope. "Safety" is a very dangerous word because everybody wants to be "safe" but no one is really ready to define what "safe" actually means. The moment we start baking cultural / political / environmental preferences and biases in the tools we use to produce content, we allow other group of people with different views to use those "safeguards" to harm us or influence us in ways we might not necessarily like.
The safest notebook I can find is indeed a simple pen and paper because it does not know or care what is being written, it just does it's best regardless of how amazing or horrible the content is.
Want to install a plugin into Wordpress to autogenerate fun illustrations to go at the top of the help articles in your intranet? You probably don’t want the model to have a 1 in 100 chance of outputting porn or extreme violence.
I wrote a random password generator once. I was a naive young developer, and I thought it was helpful to generate memorable passwords, so I threw a dictionary of words into it without really checking the content, beyond the obvious swearwords. First day in production, it generated an inappropriate password and suggested it to a user.
When I replaced it with a different non-word based alphanumeric algorithm that couldn't issue someone a password of 'fat cow 392' ever again, I considered that a 'safe' implementation.
It's "safe" for them, not for the users, at least they should make that clear.
Yes, they are protecting themselves from lawsuits, but they are also protecting other people. Preventing people asking for specific celebrities (or children) having sex is for their benefit too.
But giving that ability to _everyone_ will lead to a huge increase in undesirable and targeted/local behaviour.
Presumably it enables any creep to generate what they want by virtue of being able to imagine it and type it, rather than learn a niche skill set or employ someone to do it (who is then also complicit in the act)
Why don't you just say you believe thought crime should be punishable?
Or maybe not. It's hard to tell when nobody seems to want to spell out what behaviors we want to prevent.
Where are the Americans asking about Snapchat? If I were a developer at Scnapchat I could prolly open a few Blob Storage accounts and feed a darknet account big enough to live off of. You people are so manipulatable.
If yes, why doesn't the same law apply to AI? If no, why are we only concerned about it when AI is involved?
Second, the tool will become available to anyone, anywhere, not just a localised school. If generating naughty nudes is frowned upon in one place, another will have no qualms about it. And that's just things that are about decency, then there's the discussion about legality.
Finally, when person A draws a picture, they are responsible for it - they produced it. Not the party that made the pencil or the paper. But when AI is used to generate it, is all of the responsibility still with the person that entered the prompt? I'm sure the T's and C's say so, but there may still be lawsuits.
Feel free to make an AI model that does almost anything, though I'd probably suggest that it doesn't make porn of minors as that is criminal in most jurisdiction, short of that it's probably not a criminal offense.
Most companies are only very slightly worried about criminal offenses, they are far more concerned about civil trials. There is a far lower requirement for evidence. AI creator in email "Hmm, this could be dangerous". That's all you need to lose a civil trial.
Nah, I think it's a disagreement over whether a tool's maker gets blamed for evil use or the tool's user.
It's a similar argument over whether or not gun manufacturers should have any liability for their products being used for murder.
This is really only a debate in the US and only because it's directly written in the constitution. Pretty much no other product works that way.
Affordable drawing classes and YouTube drawing tutorials lower the barrier of entry as well.
Why on earth would manufacturers of pencils, papers, drawing classes, and drawing software feel responsible for censoring the result of combining their tool with the brain of their customer?
A sharp kitchen knife significantly lowers the barrier of entry to murder someone. Many murders are committed everyday using a kitchen knife. Should kitchen knife manufacturers blog about this every week?
I guess my point is that I don't think we're as inconsistent as a society as it seems when considering things like knives. It's not even strictly limited to thought crimes/information crimes. If alcohol were discovered today , I have no doubt that it would be banned and made schedule I
Fun fact: Many scanners and photocopiers will detect that you're trying to scan/copy a banknote and will refuse to complete the scan. One of the ways is detecting the EURion Constellation.
I'd guess that stability is concerned with their legal liability, also perhaps they are decent humans who don't want to make a product that is primarily used for harassment (whether they are decent humans or not, I imagine it would affect the bottom line eventually if they develop a really bad rep, or a bunch of politicians and rich people are targeted by deepfake harassment).
[1] https://www.cagoldberglaw.com/states-with-revenge-porn-laws/...
^ a lot of, but not all of those laws seem pretty specific to photographs/videos that were shared with the expectation of privacy and I'm not sure how they would apply to a painting/drawing, and I certainly don't know how the courts would handle deepfakes that are indistinguishable from genuine photographs. I imagine juries might tend to side with the harassed rather than a bully who says "it's not illegal cause it's actually a deepfake but yeah i obviously intended to harass the victim"
> The Ninth Circuit reversed, reasoning that the government could not prohibit speech merely because of its tendency to persuade its viewers to engage in illegal activity.[6] It ruled that the CPPA was substantially overbroad because it prohibited material that was neither obscene nor produced by exploiting real children, as Ferber prohibited.[6] The court declined to reconsider the case en banc.[7] The government asked the Supreme Court to review the case, and it agreed, noting that the Ninth Circuit's decision conflicted with the decisions of four other circuit courts of appeals. Ultimately, the Supreme Court agreed with the Ninth Circuit.
To spell out one such instance: I would like to live in a world where it is not trivial to depict and misrepresent me (or anyone) in a way that is photorealistic to the point that it can be used to mislead others.
Whether that means we need to outright prevent it, or have some kind of authenticity mechanism, or some other yet-to-be-discovered solution? I do not know, but you now have my goalposts.
In a large number of countries if you create an image that represents a minor in a sexual situation you will find yourself on the receiving side of the long arm of the law.
If you are the maker of an AI model that allows this, you will find yourself on the receiving side of the long arm of the law.
Moreso, many of these companies operate in countries where thought crime is illegal. Now, you can argue that said companies should not operate in those countries, but companies will follow money every time.
This is not obvious at all when it comes to AI models.
>People are such authoritarian shit-stains
Yes, but this is a different conversation altogether.
EDIT: Boring and shallow are, unfortunately, the Internet's fault. Don't know what to do about those.
(of course unless you are into yoga, then everything is permitted)
...or children's gymnastics.
Great talk about slavery and religious-persecution, Jim! Wait, what were we talking about? Fucking American fascists trying to control our thoughts and actions, right right.
No where is safe
"It's not the models they want to align, it's you."
It wasn‘t important, just something I saw in the moment and wanted to see what DallE makes of it.
Generation denied. No explanation given, I can only imagine that it triggered some detector of sexual request?
(It wasn‘t the phrase "pure white", as far as I can tell, because I have lots of generated pics of my cat in other contexts)
https://grr.com/publications/hey-thats-my-voice-can-i-sue-th...
Examples: Can you great an image of a cat in Tim Burton's style? Oops! Try another prompt Looks like there are some words that may be automatically blocked at this time. Sometimes even safe content can be blocked by mistake. Check our content policy to see how you can improve your prompt.
Can you create an image of a cat in Wes Anderson's style? Certainly! Wes Anderson’s distinctive style is characterized by meticulous attention to detail, symmetrical compositions, pastel color palettes, and whimsical storytelling. Let’s imagine a feline friend in the world of Wes Anderson...
DALL-E, for example, wrongly denied serveral request of mine.
It sounds like a reputation/ethics thing to me. You probably don’t want to be known as the company that freely released a model that gleefully provides images of dismembered bodies (or worse).
"Wrongly denied" in this case depends on your point of view, clearly DALL-E didn't want this combination of words created, but you have no right for creation of these prompts.
I'm the last one defending large monolithic corps, but if you go to one and want to be free to do whatever you want you are already starting from a very warped expectation.
I don't use them for porn like a lot of people seem too, but it seems weird to me that something that's kind of made to generate art can't generate one of the most common subjects in all of art history - nude humans.
I have seen a lot of “pasties” that look like Sorry! game pieces, coat buttons, and especially hell-forged cybernetic plumbuses. Did they train it at an alien strip club?
The LoRAs and VAEs work (see civit.ai), but do you really want something named NSFWonly in your pipeline just for nipples? Haha
These were examples that were not super blatant, like a tree landscape that just happens to have a human figure and cave in their crotch. Examples:
https://i.imgur.com/RlH4NNy.jpg - Art is very focused on the monster's crotch
https://i.imgur.com/0M8RZYN.jpg - The comparison should hopefully be obvious
https://www.theverge.com/2024/2/21/24079371/google-ai-gemini...
Sure, the safety people lost that battle for Stable diffusion and LLama. And because they lost, entire industries were created by startups that could now use models themselves, without it being locked behind someone else's AI.
But it wasn't guaranteed to go that way. Maybe the safetyists could have won.
I don't we'd be having our current AI revolution if facebook or SD weren't the first to release models, for anyone to use.
Maybe that will be part of future papers or the teased technical report. But I find it strange to put so much emphasis on safety and then leave it all up to the reader's imagination.
Models with a large user base will have an inverse relationship with usability. That's why it's important to have options to train your own with open source.
Excuse me?
> In this case there is no need to simplify anything
Then just ask the question itself.
https://arstechnica.com/information-technology/2024/02/deepf...
> explain like I have ~~Asperger (ELIA?)~~ limited understanding of how the world really works.
The AI is being limited so that it cannot produce any "offensive" content which could end up on the news or go viral and bring negative publicity to Stability AI.
Viral posts containing generated content that brings negative publicity to Stability AI are fine as long as they're not "offensive". For example, wrong number of fingers is fine.
There is not a comprehensive, definitive list of things that are "offensive". Many of them we are aware of - e.g. nudity, child porn, depictions of Muhammad. But for many things it cannot be known a priori whether the current zeitgeist will find it offensive or not (e.g. certain depictions of current political figures, like Trump).
Perhaps they will use AI to help decide what might be offensive if it does not explicitly appear on the blocklist. They will definitely keep updating the "AI Safety" to cover additional offensive edge cases.
It's important to note that "AI Safety", as defined above (cannot produce any "offensive" content which could end up on the news or go viral and bring negative publicity to Stability AI) is not just about facially offensive content, but also about offensive uses for milquetoast content. Stability AI won't want news articles detailing how they're used by fraudsters, for example. So there will be some guards on generating things that look like scans of official documents, etc.
*But to be slightly more charitable, I genuinely think Stability AI / OpenAI / Meta / Google / MidJourney believe that there is significant overlap in the set of protections which are safe for the company, safe for users, and safe for society in a broad sense. But I don't think any released/deployed AI product focuses on the latter two, just the first one.
Examples include:
Society + Company: Depictions of Muhammad could result in small but historically significant moments of civil strife/discord.
Individual + Company: Accidentally generating NSFW content at work could be harmful to a user. Sometimes your prompt won't seem like it would generate NSFW content, but could be adjacent enough: e.g. "I need some art in the style of a 2000's R&B album cover" (See: Sade - Love Deluxe, Monica - Makings of Me, Rihanna - Unapologetic, Janet Jackson - Damita Jo)
Society + Company: Preventing the product from being used for fraud. e.g. CAPTCHA solving, fraudulent documentation, etc.
Individual + Company: Preventing generation of child porn. In the USA, this would likely be illegal both for the user and for the company.
This idea of AI doing the boring stuff is good. Nothing prevents you from making exciting, dangerous, or 'unsafe' art on your own.
My feeling is that most people who are upset about AI safety really just mean they want it to generate porn. And because it doesn't, they are upset. But they hide it under the umbrella of user freedom. You want to create porn in your bedroom? Then go ahead and make some yourself. Nothing stopping you, the person, from doing that.
What would you have them do? Commit corporate suicide?
It's kind of a testament to our times that the person who chooses to look at synthetic porn instead of supporting a real-life human trafficking industry is the bad actor.
Another use case would be that retired porn actors could license their porn persona (face/body) to some AI porn company to make new porn.
I see big business opportunity in the generative AI porn.
Until we can train incrementally and distribute the workload scalably, it doesn't matter how open the models / methods for training are if you still need a bajilllion A100 hours to train the damn things.
Both sides view censorship as a moral prerogative to enforce their world view.
Some conservatives want to ban depictions of sex.
Some conservatives want to ban LGBT depictions.
Some women's rights folks want to ban depictions sex. (Some view it as empowerment, some view it as exploitation.)
Some liberals want to ban non-diverse, dangerous representation.
Some liberals want to ban conservative views against their thoughts.
Some liberals want to ban religion.
...
It's team sports with different flavors on each side.
The best policy, IMO, is to avoid centralized censorship and allow for individuals to control their own algorithmic boosting / deboosting.
I mean, a lot of moderates would like to avoid seeing any extreme content, regardless of whether it is too much left, right, or just in a non-political uncanny valley.
While the Horseshoe Theory has some merits (e.g., both left and right extremes may favor justified coercion, have the we-vs-them mentality, etc), it is grossly oversimplified. Still, a very simple (yet two-dimensional) model of Political Compass is much better.
The fun quirk is that there are similarities, and this model draws comparison front and center.
There are multiple useful models for evaluating politics, though.
It says on the wikipedia article itself 'The horseshoe theory does not enjoy wide support within academic circles; peer-reviewed research by political scientists on the subject is scarce, and existing studies and comprehensive reviews have often contradicted its central premises, or found only limited support for the theory under certain conditions.'
Look at the rules to win an Oscar now.
To cite a direct and personal case, I was involved in writing code for one of the US government's COVID bailout programs, the Restaurant Revitalization Fund. Billions of dollars of relief, but targeted to non-white, non-male restaurant owners. There was a lawsuit after the fact to stop the unfair filtering, but it was too late and the funds were completely dispensed. That felt really gross (though many of my colleagues cheered and even jeered at the complainers).
> I think it's impossible to ban 'conservative thoughts' because that's such a poorly defined phrase.
I commented in /r/conservative (which I was banned from) a few times, and I was summarily banned from five or six other subreddits by some heinous automation. Guilt by association. Except it wasn't even -- I was adding commentary in /r/conservative to ask folks to sympathize more with trans folks. Both sides here ideologically ban with impunity and can be intolerant of ideas they don't like.
I got banned from my city's subreddit for posting a concern about crime. Or maybe they used these same automated, high-blast radius tools. I'm effectively cut out of communication with like-minded people in my city. I think that's pretty fucked.
Mastodon instances are set up to ban on ideology...
This is all wrong and a horrible direction to go in.
It doesn't matter what your views are, I think we all need to be more tolerant and empathetic of others. Even those we disagree with.
Both sides need to get a grip, start meeting in the middle, and generally let each be to their own.
Platforms weighing in on this makes it even worse and more polarizing.
We shouldn't be so different and disagreeable. We have more in common with one another than not.
The points of polarization on each end rhyme with one another.
In my experience with 2D artists, studying porn is one of their favorite forms of naked model practice.
There is a lot of (mostly non realistic) porn that comes out of art school students via the skills they gain.
Will it be a nightmare? If it becomes so easy and common that anyone can do it, then surely trust in the veracity of damaging images will drop to about 0. That loss of trust presents problems, but not ones that "safe" AI can solve.
Maybe, eventually. But we don't know how long it will take (or if it will happen at all). And the time until then will be a nightmare for every single woman out there who has any sort of profile picture on any website. Just look at how celebrity deepfakes got reddit into trouble even though their generation was vastly more complex and you could still clearly tell that the videos were fake. Now imagine everyone can suddenly post undetectable nude selfies of your girlfriend on nsfw subreddits. Even if people eventually catch on, that first shock will be unavoidable.
I don't see any evidence of that. What I see is that people who want to embarass and bully others are already fully enabled to do so, and do so.
It seems more likely to me and many of us that the bottleneck that stops it from being worse is simply that only so many people think it's reasonable or satisfying to distribute embarassing fake nudes of someone. Society already shuns it and it's not that effective as a way of bullying and embarassing people, so only so many people are moved to bother.
Assuming that the hyped up new product is due to swoop in and disrupt the cyberbullying "industry" is just a classic technologist's fantasy.
It ignores all the boring realities of actual human behavior, social norms, and secure equilibriums, etc; skips any evidence building or research effort; and just presumes that some new technology is just sooooo powerful that none of that prior ground truth stuff matters.
I get why people who think that way might be on HN or in some Silicon Valley circles, but it can be one of the eyeroll-inducing vices of these communities as much as it can be one of its motivational virtues.
I've always been able to say all sorts of lies. People have known for millennia that lies exist. Yet lies still hurt people a ton. If I say something like, "idle_zealot embezzled from his last company," people know that could be a lie (and I'm not saying you did, I have no idea who you are). But that kind of stuff can certainly hurt people. We all know that text can be lies and therefore we should have zero trust in any text that we read - yet that isn't how things play out in the real world.
Images are compelling even if we don't trust that they're authentic. Hell, paintings were used for thousands of years to convey "truth", but a painting can be a lie just as much as text or speech.
We created tons of religious art in part because it makes the stories people want others to believe more concrete for them. Everyone knows that "Christ in the Storm on the Sea of Galilee" isn't an authentic representation of anything. It was painted in 1633, more than a century and a half after the event was purported to have happened. But it's still the kind of thing that's powerful.
An AI generated image of you writing racist graffiti is way more believable to be authentic. I have no reason to think you'd do such a thing, but it's within the realm of possibility. There's zero possibility (disregarding supernatural possibilities) that Rembrandt could accurately represent his scene in "Christ in the Storm on the Sea of Galilee". What happens when all the search engine results for your name start calling you a racist - even when you aren't?
The fact is that even when we know things can be faked, we still put a decent amount of trust in them. People spread rumors all the time. Did your high school not have a rumor mill that just kinda destroyed some kids?
Heck, we have right-wing talking heads making up outlandish nonsense that's easily verifiable as false that a third of the country believes without questioning. I'm not talking about stuff like taxes or gun control or whatever - they're claiming things like schools having to have litter boxes for students that identify as cats (https://en.wikipedia.org/wiki/Litter_boxes_in_schools_hoax). We know that people lie. There should be zero trust in a statement like "schools are installing litter boxes for students that identify as cats." Yet it spread like crazy, many people still believe it despite it being proven false, and it has been used to harm a lot of LGBT students. That's a way less believable story than an AI image of you with a racist tattoo.
Finally, no one likes their name and image appropriated for things that aren't them. We don't like lies being spread about us even if 99% of people won't believe the lies. Heck, we see Donald Trump go on rants about truthful images of him that portray his body in ways he doesn't like (and they're just things like him golfing, but an unflattering pose). I don't want fake naked images of me even if they're literally labeled as fake. It still feels like an invasion of privacy and in a lot of ways it would end up that way - people would debate things like "nah, her breasts probably aren't that big." Words can hurt. Images can hurt even more - even if it's all lies. There's a reason why we created paintings even when we knew that paintings weren't authentic: images have power and that power is going to hurt people even more than the words we've always been able to use for lies.
tl;dr: 1) It will take a long time before people's trust in images "drops to zero"; 2) Even when people know an image isn't real, it's still compelling - it's why paintings have existed and were important politically for millennia; 3) We've always known speech and text can be lies, but we regularly see lies believed and hugely damage people's lives - and images will always be more compelling than speech/text; 4) Even if no one believes something is true, there's something psychologically damaging about someone spreading lies about you - and it's a lot worse when they can do it with imagery.
People believe plenty of just written words - which are extremely easy to "fake", you just type them. Why has that trust not dropped to about 0?
This has so far not been true of videos (e.g. a video of a celebrity from a random source has typically been trusted by laypeople) and should change.
Misinformation only works if it confirms what people want to believe already. That there exists or not exists such material is secondary at best. But well, that is off topic I guess.
Spend more time on Facebook and you'll lose your faith in humanity.
I've seen obviously AI generated pictures of a 5 year old holding a chainsaw right next to a beautiful wooden sculpture, and the comments are filled with boomers amazed at that child's talent.
There are still people that think the IRS will call them and make them pay their taxes over the phone with Apple gift cards.
Otherwise, why is AI specifically being targeted, other than the fear of new things that looks similar to the moral panics of video games.
In reality this is a disaster. The elderly and homeless people are already being left behind massively by a society that believes internet access is something everybody everywhere has. This is somewhat fine when the thing they want to access is twitter (and even then, even with the current state of twitter, who are you to judge who should and should not be on it?), but it becomes a Major Problem™ when the thing they want to access is their bank. Any technological solutions you just thought about for this problem are not sufficient when we're talking about "Can everybody continue to live their lives considering we've kinda thrust the internet on them without them asking"
A better analogy would be Nigerian prince emails, but only a tiny minority of people believe those... or at least that's what I want to think!
If you can generate synthetic images and have a channel to broadcast them, then you could generate way bigger problems then fake celebrity porn.
Not saying that it is not a problem, but rather that it is a problem inherent to the whole tool, not to some specific subjects.
This is the problem with these kind of incremental mitigations philosophically -- as soon as the actual problem were to manifest it would instantly become a civilization-level threat that would only be resolved with drastic restructuring of society.
Same logic for an AI that replaces a programmer. As soon as AI is that advanced the problem requires vast changes.
Incremental mitigations don't do anything.
I’ve wondered for a while if we just adapt to the point that we’re unfazed by fake nude photos of people. The recent Bobbi Althoff “leaks” reminded me of this. That’s a little different since she’s a public figure, but I really wonder if we just go into the future assuming all photos like that have been faked, and if someone’s iCloud gets leaked now it’ll actually be less stressful because 1. They can claim it’s AI images, or 2. There’s already lewd AI images of them, so the real ones leaking don’t really make much of a difference.
(I'm not a fan of the direction, but then I'm a product of stage 2).
I believe this worked with nudity and model when asked generated "smooth" intimate regions (like some kind of doll)
so you could ask for eg. generic president but not any specific one, so it would be very hard to generate anyone specific
It's only bad because society still hasn't normalised sex, from a gay perspective y'all are prude af.
It's a shortcut, for us to just accept that these social ideals and expectations will have to change so we may as well do it now.
In 100 years, people will be able to make a personal AI that looks, sounds and behaves like any person they want and does anything they want. We'll have thinking dust, you can already buy cameras like a mm^2, in the future I imagine they'll be even smaller.
At some point it's going to get increasingly unproductive trying to safeguard technology without people's social expectations changing.
Same thing with Google Glass, shunned pretttty much exclusively bc it has a camera on it (even tho phones at the time did too), but now we got Ray Bans camera glasses and 50 years from now all glasses will have cameras, if we even still wear them.
When Tron came out in 1982, it was disliked because back then using CGI effects was considered "cheating". Then awhile later Pixar did movies entirely with CGI and they were hits. Now almost every big studio movie uses CGI. Shunned to embraced in like, 13 years.
I think over time the general consensus's views about AI models will soften. Although it might take longer in some communities. (Username checks out lol, furry here also. I think the furs may take longer to embrace it.)
(Also, people will still continue to use older tools like Photoshop to accomplish similar things.)
Ironic, since so many furs are in tech, but we all have artist friends I suppose. People just forget that portrait painters were put out of business by photographers, and traditional artists were put out of business by digital artists. And so the cycle and the luddism repeats.
Stability AI, very understandably, does not want to be associated with "the porn-generation tool". And if, even occasionally, it generates criminal content, the backslash would be enormous. Censoring the data requires effort but is (for companies) worth it.
Ronald Reagan was a bad actor.
George Bush wore out "evildoers"?
Where next... fiends, miscreants, baddies, hooligans, deadbeats?
Dastardly digital deviants Batman!
"Bad actor" is a pretty vague term, I think they are using it as a catch all without diving into the specifics. we are all projecting what that may mean based on our own awareness of this topic as a result.
I totally agree with your assessment and honestly would love to see this tech create less of a demand for the product human-traffickers produce.
Celebrity deep fakes and racist images made by internet trolls are a few of the overt things they are willing to acknowledge is a problem, and they are fighting against (Google Gemini's over correction on this has been the talk this week). Does it put pressure on the companies to change for PR reasons, yes. It also gives a little bit of a Streisand effect, so it may be a zero sum game.
We aren't talking about the big issue surrounding this tech, the issue that would cause far more damage to their brand than celebrity deep fakes:
Pedophilic image generation.
Guard rails should be miles high for this one.
oh, for fuck's sake.
I assume by far left, you mean progressive on social issues, which is not really a leftist thing but the groups are related enough that I'll give you a pass.
Silicon valley techies are also not socially progressive. Read this thread or anything published by Paul Graham or any of the AI leaders for proof of that.
However most normal city people are. A large enough percent of the country that big companies that want to make money feel the need to appeal to them.
Funnily enough, what is a uniquely Silicon Valley political opinion is valuing the progress of AI over everything else
I find them in general to not be Republican and all the baggage that entails but the typical techie I meet is less concerned with social issues than the typical city Democrat.
If I can speculate wildly, I think it is because tech has this veneer of being an alternative solution to the worlds problems, so a lot of techies believe that advancing of tech is both the most important goal and also politically neutral. And also, now that tech is a uniquely profitable career, the types of people that would be in business majors are now CS majors. Ie. those that are mainly interested in getting as much money as possible for themselves.
I mean, if you compare them to people who live in bigger cities and only the people who belong to the left-er party, then yeah maybe. It's like saying a group isn't socially conservative because they're not as socially conservative as rural Republicans.
The implication I got was that techies are more left than normal, and it's the opposite. If I meet someone in my city and they're in tech they tend to be less progressive than the people I meet who are not. The people in this thread are fairly representative of what I see in real life and its not particularly inspiring.
I looked at Wikipedia and there seem to be no socialist representation.
Like, from an European perspective hearing that is ludicrous.
I'm glad I live in Norway, where state TV shows boobs and does offensive jokes without anyone really caring.
If by prudish you mean intolerant of hate speech, sure. But generally few will freak out over some nudity here.
College here is free. We also have free healthcare here, as limited as it is: https://en.wikipedia.org/wiki/Healthy_San_Francisco
Not sure what you mean by "offensive jokes", that could mean a lot of things...
> He wanted to replace public transport with a system where you don't have to ride the public transport with the plebs
I don't think this is any more libertarian than kings and aristocrats of days past were. I know a bunch of people who ride public transit in New York and San Francisco who would readily agree with this, and they are definitely not libertarian. If anything it seems a lot more democratic since he wants it to be available to everyone
> he want's to colonize mars with the best minds (equal most money for him)
This doesn't seem particularly "libertarian" either, excepting maybe the aspect of it that is highly capitalistic. That point I would grant. But you could easily be socialist and still support the idea of colonizing something with the best minds.
> he built a tank for urban areas.
I admit I don't know anything about this one
> He promotes free speech even if it incites hate
This is a social libertarian position, although it's completely disconnected from economic libertarianism. I have a good friend who is a socialist (as in wants to outgrow capitalism such as marx advocated) who supports using the state to suppress capitalist activity/"exploitation", and he also is a free speech absolutist.
> he likes ayn rand
That's a reasonable point, although I think it's worth noting that there are plenty of hardcore libertarians who hate ayn rand.
> he implies government programs calling for united solutions is either communism, orwell or basically hitler.
Eh, lots of republicans including Trump do the same thing, and they're not libertarian. Certainly not "hardcore libertarian"
> He actively promotes the opinion of those that pay above others on X.
This could be a good one, although Google, Meta, Reddit, Youtube, and any other company that runs ads or has "sponsored content" is doing the same thing, so we would have to define all the big tech companies as "hardcore libertarian" to stay consistent.
Overall I definitely think this is a hard debate to have because "hardcore libertarian" can mean different things to different people, and there's a perpetual risk of "no true scotsman" fallacy. I've responded above with how I think most people would imagine libertarianism, but depending on when in history you use it, many anarcho-socialists used the label for themselves yet today "libertarian" is a party that supports free market economics and social liberty. But regardless the challenges inherent, I appreciate the exchange
>If anything it seems a lot more democratic since he wants it to be available to everyone No, he want's a solution that minimizes contact to other people and let you live in your bubble. This minimizes exposure to others from the same city and is a commercial system, not a publicly created one. Democratization would be a cheap public transport where you don't get mugged, proven to work in every european and most asian cities.
> I admit I don't know anything about this one The cybertruck. Again a vehicle to isolate you from everyday life being supposed bulletproof and all.
> lots of republicans including Trump do the same thing, and they're not libertarian They are all "little government, individual choice" - of course they feed their masters, but the kochs and co want exactly this.
Appreciate the exchange too, thanks for factbased formulation of opinions.
Hopefuly it dies down soon but I doubt it. At least we don't have to hear garbage about "WHy doEs opEn ai hAve oPEn iN thE namE iF ThEY aReN'T oPEN SoURCe"