Copilot stops working on code that contains hardcoded banned words from GitHub (2023)
github.com
github.com
Not to mention, there’s apparently some research saying code with swear words has higher quality, so if AI causes some decline there, we now know why it is https://www.reddit.com/r/programming/comments/110mj6p/open_s...
If this LLM stuff is as important as it is made out to be (and I think it is), it is absolutely crucial that it isn’t controlled by just a bunch of large corporations and tech oligarchs. A world where everybody needs LLM’s and the only source is from gigantic tech companies would be incredibly distopian.
Plus I seriously doubt the true innovation on these things will happen until the “little guy” can train and infer their own models. Right now these things are playing it way to safe. We should be reading articles about people using home built LLMs to do crazy shit that sticks it to “The Man”. That’s how transformative technology changes things. It challenges the status quo. The only “quo” that is being challenged right now is some boring mega tech company is peeing in some other mega tech company’s cheerios. Yawn. Wake me up when this technology threatens the entire fucking system (and not fake “AGI will take over all white collar jobs”… that is just corporate propaganda).
… I mean remember Napster and all those file sharing companies? Or the million iterations of Pirate Bay? Where is the LLM equivalent of that? Where is the “Linux” of LLM’s that freaks out all the tech companies? Or dark web of LLM’s that attract the eye of every three letter agency in the world? Where is the revolution? It’s just a bunch of huge tech companies safely jerking each other off wearing three layers of protection and their corporate lawyers on speed dial. How completely boring.
(cl-defstruct person gender)
(make-person :gender "m<|>")
with cursor at <|> does elicit "male" completion. Yay for normalcy?(Though honestly, I didn't notice this earlier - Copilot tends to hang for me too often in all kinds of files for me to identify these stopwords.)
(EDIT: gotta admit though, this is hilarious: https://github.com/orgs/community/discussions/72603#discussi...)
From my testing, I see that one still shuts down the Copilot completely. It may mean "late" in French, but it's also often used (at least where I live) to mark SR/Slow Release versions of drugs. Apparently, now, even writing software for pharmacies is immoral and should be blocked... :D
And when the decision-maker can also steer the narrative...say...by mobilizing downvotes...The outcome is predictable. :-))
Companies that are ironically filled with privileged people who "Gotta do something", since they, while sort-of well meaning, are sheltered and disconnected from actual social struggles in real life.
A few months ago, I visited San Jose for a wedding. When I picked up my rental car at the airport, the only options were electric, even though I had specified that I wanted an ICE vehicle tat the time I made the reservation. During the four day trip I wound up visiting five different charging stations (some of which slow charged, so weren't able to replenish the battery in the time available), and I had to install three different apps. I still have like $20 of unused credit between them. I spent several hours waiting for the car to charge, not to mention making major detours looking for a fast-charging station. If I were to guess a part of the world you'd expect to find the best possible electric vehicle charging infrastructure, San Jose wouldn't be far off the mark. But my trip wound up being dominated by range anxiety.
I drive a PHEV (plug-in hybrid) for my commute from home to work and I love it. Electric is the future, and it's great that electric vehicles and charging stations are becoming more common. But renting an electric car in a strange city today is about the worst possible scenario for a short range vehicle. You don't know where the charging stations are, the charging stations require different apps, your hotel might not have a charger, and so on. The people making decisions at car rental companies should know this!
They are not being randomly paranoid. Even if they did not have this fear, they would have rapidly developed it. We've all read the articles by muckraking journalists that take something an AI said and basically deliberately writes clickbait about how stupid or evil or worthless or whatever the AI is, even if the journalist had to filter through hundreds of replies (or, implicitly, by waiting for the dumbest stuff to rise to the top of social media, thousands or millions of replies) to get it. We've also read the articles where in someone uses the "fancy autocompleter", feeds it the moral equivalent of "Hey, how do you think you AIs will be taking over the world in five years?" and then is shocked, shocked at the "fancy autocompleter" filling in the yarn they are clearly asking for, and go running to either the media, or in particularly pathological cases, the academic literature making wild claims.
(I do not believe that "fancy autocompleter" is a complete description of LLMs, but in this particular case, it isn't a completely inaccurate mental model either. It shouldn't be a surprise that when you prompt it with X, you get more X.)
As a result the AIs are very heavily tuned to some combination of the political beliefs of the company writing them and the political beliefs dominant in the media coverage they are worried about, so they won't get very negative stories written about them. For this purpose, I'm taking the broadest possible definition of "political", not just "American politics in 202x", but the full range of "beliefs that not everyone agrees on and are things people are willing to exert some degree of power over". The AI companies have to take a stand, because taking a stand at least means someone can be on their side... if they just let the chips fall where they may they'll anger everyone because everyone can get the AI to say things that they in particular disagree with and they'll find themselves without friends. Unsurprisingly, the AI companies have been aligning their models with what they perceived to be the largest, most powerful political beliefs in their vicinity.
To be honest when I read them talking about "AI safety" I know they want me to be thinking "ensuring the AI doesn't take over the world or tell people to commit self harm" but what I see is them spending a lot of effort to politically align their AIs, with all that entails.
Oh it's certainly not by many people. It's fine with them if you are censoring what they don't like.
It took me a long time to realize the "free speech" advocates of the 60s-80s were long gone from the left.
You clearly have a moral spine because you're standing up and asking questions about what is right.
LLMs can't do anything like that so they need to be constrained by more direct methods in order to not appear to be evil
There's still career-ending possibilities from saying things even in the overtly antiwoke government: https://www.independent.co.uk/news/world/americas/us-politic... ; people are not going to trust an LLM which might say such things on their behalf.
Not to mention questions like "who is liable if an LLM libels someone".
https://www.theguardian.com/technology/2024/nov/04/google-me...
Because of the obvious PR implications of having a program one's company wrote spewing controversial takes. That's what it boils down to - and it's entirely reasonable.
Personally, I wish these things could have a configurable censorship setting. Everyone has different things that get under their skin, after all (and this would satisfy both the pro-censor and pro-uncensored groups). It's a good argument for self hosting, too, because those can be filtered to your own sensitivities.
That would help with cases where the censorship is just dead wrong. A friend was working with one of the coding ones in VS Code, and expressed his frustration that as soon as the codebase included the standard acronym for "Highest Occupied Molecular Orbital" (HOMO) it just refused any further completion. We both guessed the censor was catching it as a false positive for the slur.
That's the right answer. And it's not like this is a potential risk that is only being theorized about. Microsoft already has a very hands-on experience with disasters of this exact nature:
There is a word for this. It is called the Scunthorpe problem. Named after the incident in which the residents of the Town Scunthorpe could not register for an AOL account because AOL had an obscenity filter that did not allow the Town name.
It has been a problem since 1996 and still causes problems.
And that's why we can't have nice things.
It's not. Most people do support censorship. They just don't admit that.
It's somewhat similar to the laws some places have against providing free alcohol. Alcohol is still legal and abuse still happens. However, at least requiring people to spend money provides some friction to prevent things from escalating too much.
There are a lot of things I can say as a citizen that would get me fired from my job, or at least a talking-to by someone in management.
At the moment at least these LLMs are mainly hosted services branded by the companies that trained and/or operate them. Having a Microsoft-branded LLM say something that Microsoft as a corporation doesn't want said is something they will try to control.
That's also different from thinking that all LLMs should be censored. You can train or run your own with different priorities if you wish. It's like how there's a lot of media out there that you can consume or create yourself perfectly legally that isn't sold at Walmart.
Simple: the people who are very pro free speech (i.e. "censorship is evil"), and the people who want to censor LLMs are distinct groups (though both groups are vocal).
If I make a word processor, it doesn't need any stance on the Israel/Palestine conflict. It's just a word processor.
But if I make an LLM, and you prompt it to tell you about the Israel/Palestine conflict? The output will be deeply political, and if it refuses to answer that will also be political
The technology industry does not know what to do because unlike industries like journalism and publishing who are used to engaging with politics, a lot of norms, power structures and people in tech think we're still in the 1990s making word processors, no politics here.
Almost every "everything goes" type forum becomes undesirable to almost everyone for a variety of reasons.
Users might complain about it, but they also don't want "no censorship" even if that's what they say.
I didn't intend to imply that, but I see how my wording was unclear.
I mean that LLMs don't appear to be up for these censorship-like tasks. The evidence being that a highly visible team using LLMs uses much older tech for a highly visible function. It's useful to know the limits of tech, especially novel tech, and this use case appears to be one.
It feels like nobody's working to improve the autocomplete/copilot experience, everyone's focused on the "chat with the code and get AI to make all the changes for you" instead of "I know what I'm doing, just predict what I'm about to type and save me the effort of typing it out".
But lets not pretend its for the benefit of users. If a company could release an unmoderated model without real risk to themselves then they should do so.
Copilot stops working on gender related subjects
E.g. a system design interview question about building an app to manage all your drug shipments and nuclear bombs.
• https://www.haaretz.com/2010-01-20/ty-article/news-site-call...
• https://en.wikipedia.org/wiki/Scunthorpe_problem
• My dad had a story about an all-staff memo about an "African-American tie event".
• I had warnings from Apple about using "Knopf" in a description, which can only have come from the English word "knob" being literally (and inappropriately) translated into German from an English-language bad word filter, as "Knopf" isn't at all rude in German.
But not this: https://skeptics.stackexchange.com/questions/31343/did-a-sur...
But I'm not a huge fan of relying on online coding tools. Has anyone tried running Deepseek locally and use it for coding?
;; Created: May 1987
Guess I'm one of today's lucky 10000 :)I think one concrete thing is writing a comment about a function, then expecting that function below the comment, but instead you get more comments.
That is to say, the release version of ChatGPT GPT4o seemed much better than the version today. This does not apply to the API.
I prefer to just write code myself
If I used copilot, it would prevent the doctors from recording the patient's sex?
This is joke, actually I used it for that. But not for cool nuclear stuff that you see in the movies.
And if you don't get the joke. It is about o3-mini model card having couple of pages about how they prevent it answering some nuclear and radiation questions.
> The Apple software is not intended for use in the operation of nuclear facilities, aircraft navigation or communication systems, air traffic control systems, life support machines or other equipment in which the failure of the Apple software could lead to death, personal injury, or severe physical or environmental damage.
sigh.
My guy, most major companies struggle to get a quarterly plan in place for what they're going to do by the start of the following quarter. There may be some department focused on lobbying and political risk that cares about this stuff, but there are no product and engineering teams at any company that have "contingency plans" just laying around to be activated for random shit like this. Nobody has time for that.
I haven’t seen any recent OKR or KPI that was able to justify or even quantify the value of a path in history that didn’t occur.