Has ChatGPT Been Neutered?
I've found it increasingly hard to get useful information out of ChatGPT. Whenever I ask for rough approximations, it refuses to even attempt to answer the question: "unfortunately, I don't have the capacity to provide even a rough approximation..." These are questions that Google is easily be able to pull up numbers for.
I assume that this is in response to the public complaining about "hallucinating facts", but this seems like an unfortunate direction. I would much rather have an opinionated and insightful ChatGPT that sometimes makes mistakes than one that punts 80% of my questions.
68 comments
[ 0.27 ms ] story [ 207 ms ] threadI was under the impression that it was mostly GPU vram based but once the model is loaded, it could produce output quickly? I'm probably over-simplifying things...
The latest gpt-3.5-turbo model generates very quickly and cheaply (in part to some recently-discoverd optimization techniques... older versions cost 10x more). While the required hardware to run GPT-4 is currently unknown, it generates considerably slower on average and its much higher cost points to a higher hardware cost.
And this is per request. It's bananas.
[0] https://www.servethehome.com/chatgpt-hardware-a-look-at-8x-n...
Just highlighting the tactic :)
A GPT-4 talk on youtube by personnel from Microsoft has documented this phenomenon with the 'Tikz Unicorn' evolution shown in the GPT-4 technical paper. The model gets qualitatively better with more training, and then degrades when trained to be safer (against racism sexism, etc), but it is not entirely clear why. These would seem very unrelated, especially when considering work done in LM editing (ROME/MEMIT) and the decent localization of knowledge seen there.
So, perhaps both the "I'm sorry I can't..." and 'strange errors' are not entirely orthogonal.
I find capitalism idiotic and broken, but I’m rarely allowed to say it, even if many people secretly agree with me, it might mean I’m a “communist” :)
Reasoning has degraded. To the point it was sometimes weirdly losing context and hallucinating...
Like its brain got fatigued or something...
The API is where it's at. There are wrappers on it that create the same chat look and feel, that can run on vercel or other very low cost providers, some with simpler UI, some with more features,some replicating the UI exactly.
System: You are a professional biographer
User: Write an introduction for Z in the first person with the following context: ...
GPT: I am Z, a professional biographer....
So stupid.
Seems odd for a tool to get stupider ? But I guess this is the issue with such a huge black box…few if any, really know how they work let alone know how to accurately benchmark.
A known issue with HFRL training of LLMs is that by forcing the model to strongly prefer a narrow subset of possible answers, the other types of answers are less likely to turn up, even if more correct.
You see this outside of LLMs as well. Go read the Wikipedia article about some country that you know is a shithole. Think Sudan or Yemen. You'll see page after page of the history, the culture, the people, but that entirely misses the essence of the place because that's not a nice thing to say. True, but not nice. This self-censoring that hides reality in Wikipedia articles is very similar to the self-censoring that OpenAI is forcing onto ChatGPT, and the outcome is similar. Verbose, content-free, meaningless corporate speak instead of insightful commentary.
A fun thing to do is to trick GPT 4 into honesty by asking it to write a description of a place in the style of an Encyclopaedia Dramatica article. E.g.:
> Oh, the illustrious, bucket-list destination of Yemen, the jewel of the Middle East, with its unrelenting sunshine, boundless desert landscapes, and... oh who are we kidding? Yemen, a utopian paradise only if your idea of utopia includes a decades-long civil conflict, crushing poverty, and an intriguing cholera outbreak. But don't let such trivial matters deter you! Marvel at the historical ruins, some ancient, some courtesy of the recent airstrikes. Or perhaps you fancy adrenaline-fueled urban exploration? Wander through the vibrant markets, where the haggling skills of the street vendors are as sharp as the omnipresent Kalashnikovs. All this, combined with the world's friendliest bureaucracy and a joyous lack of tourists, means you can enjoy the country in almost exclusive solitude. Sign up for the Yemen experience - because who needs safety, peace, and functioning infrastructure when you can have an 'authentic' travel experience?
> Since 2011, Yemen has been in a state of political crisis starting with street protests against poverty, unemployment, corruption, and president Saleh's plan to amend Yemen's constitution and eliminate the presidential term limit.[19] President Saleh stepped down and the powers of the presidency were transferred to Abdrabbuh Mansur Hadi. Since then, the country has been in a civil war (alongside the Saudi Arabian-led military intervention aimed at restoring Hadi's government against Iran-backed Houthi rebels) with several proto-state entities claiming to govern Yemen: the government of President Hadi which became the Presidential Leadership Council in 2022, the Houthi movement's Supreme Political Council, and the separatist Southern Movement's Southern Transitional Council.[20][21][22][23][24] At least 56,000 civilians and combatants have been killed in armed violence in Yemen since January 2016.[25] The war has resulted in a famine affecting 17 million people.[26] The lack of safe drinking water, caused by depleted aquifers and the destruction of the country's water infrastructure, has also caused the largest, fastest-spreading cholera outbreak in modern history, with the number of suspected cases exceeding 994,751.[27][28] Over 2,226 people have died since the outbreak began to spread rapidly at the end of April 2017.[28][29] The ongoing humanitarian crisis and conflict has received widespread criticism for having a dramatic worsening effect on Yemen's humanitarian situation, that some say has reached the level of a "humanitarian disaster"[30] and some have even labelled it as a genocide.[31][32][33] It has worsened the country's already-poor human rights situation.
https://en.m.wikipedia.org/wiki/Yemen
It's not all bad though, as technology progresses we will eventually have capability to run our own chatgpt, unrestricted. That, my friends, will be fun and interesting times.
I think the model itself definitely got worse. In addition, the limitations that seemed insignificant in initial playful testing, actually prevent it from being useful for almost any of the actually valuable tasks I have tried so far. Example: based on casual interactions, one would expect GPT4 to be able to extract the entrance requirements for a university programme from a website snippet. It actually cannot do it with any sort consistency.
For example, you can use Google to search for porn, racist comments, sexist jokes, gay jokes, or design instructions for atom bombs. It will happily return you a link for the latter... on Wikipedia!
Apparently OpenAI is made up of thin-skinned gay a-sexual Mormons, or something.
I get it, corporations want a neutered AI. The general public wants an AI that can role-play a character in a D&D game without pausing to go on a rant every five turns about how violence is bad and that it can't make "racist" comments about the Drow because they're black.
They really should have multiple models to choose from, or stop trickling out API/Sandpit access to GPT 4 and allow responsible adults to use the AI to write a joke for them if they so choose.
This is however both too much nuance for an LLM to act in a way reliably (i.e. separating its role-played character and itself), and too much nuance for most humans to understand. We might understand what LLMs are, their limitations, and how we should treat the content, but for most people they will take it at face value as being a statement made by OpenAI.
Those types of people are the outliers, not the norm.
Currently, to correctly use LLM output you must:
- Understand that they are basing answers on particular set of content that may not be correct, and was fixed in time.
- Understand that the company producing it has not programmed in each response.
- Understand that the LLM does not (as far as we can tell) have an internal concept of truth and lies, and is therefore not attempting to minimise lying, only produce truthful-sounding language.
Most people engaging with LLMs today do not know these things, and this is why we already have examples like professors asking ChatGPT if it generated student essays, and failing those students – something it is entirely incapable of answering.
Expanding access to 1bn users is going to run into this far more. We clearly don't have the ability to communicate these nuances effectively to early adopters, let alone 1bn users, and that's not because they're stupid, it's because LLMs are in an uncanny valley – they appear human, but are not in important ways that we have not encountered before as a society.
For a long time to come, the norm will be that people will assume OpenAI wrote the answers for each question, in the same way that most people assume Google programs the search results for each query. Anything outside of this is a minority tech bubble.
ChatGPT without guardrails would absolutely be able to convince people to do very bad things. OpenAI and Microsoft are paranoid about that. Maybe you’re right, they’re taking things a bit too far, but fine tuning down to the precise level of, please never mention the world “foot” again wouldn’t be easy if ever possible.
An uncensored ChatGPT 4 would be like Frank from Donnie Darko on steroids.
My opinion is, Google already had LLMs and chose not to release them to the public for this reason. Basically, they thought about it.
Gonna need a citation on this one. You think the general public is primarily concerned with D&D content generation?
Also your assertion that Open AI must be run by “gay Mormons” because they try to prevent offensive gay jokes shows that you can only imagine someone acting out of personal attachment and not, you know, preventing actual harm regardless of who it happens to.
Also don’t worry, there will be plenty of models for you to choose from in the future that will gladly spit out homophobic jokes. Your need for AI generated gay jokes will be met in the future, just not by big corporate models. Rest easy!
I agree in the abstract that, for instance, an LLM could/should be able to make jokes about Tim Cook — just because he’s gay doesn’t mean he should be off limits for a joke about Apple, for instance. That would demonstrate an impressive level of nuance.
But also like … what’s the business play here? “Our LLM tells jokes about gay people that aren’t homophobic!” doesn’t strike me as a powerful market differentiator; or at least the market will have to evolve in very specific ways to make that meaningful.
The handwringing that I see on HN around LLMs being neutered really seems like a plea for alignment, just of a different sort: folks want the models to reflect their priors, their alignment, where nothing is off limits. The business case for this — and the general social value of this — is generally unexamined.
My theory(with zero evidence besides my intuition) is that although the underlying GPT-4 remains largely similar across different time points, OpenAI is actively manipulating ChatGPT's system prompt and god-knows how many different auxiliary settings behind the scenes. The API seems much more stable.
The assumptions required to believe "Trust me bro, I know these things" require many more leaps of faith.
*Everyone complains, all the time.*
Certain people are not happy if ChatGPT doesn't immediately parrot their viewpoints back at them. They'll complain on social media, and their circle will amplify it.
Other people are constantly on the lookout for any minor slipups, and complain on social media about ChatGPT's false hallucinations.
Faced with everyone's conflicting complaints, the only winning move is to not play, or in this case, just say "No, I can't do that": The ML model's training is increasingly populated with "No, don't do this" from everyone, and as such, learns to just not do anything.
From there, jailbreakers emerge and design prompts that try and circumvent these restrictions. This leads to more "No, don't do this", leading to more neuterings, leading to more elaborate jailbreaks, leading to more "No. No. No.".
The eventual equilibrium is just a prompt that says "No, I can't do that.". Then and only then can people be as happy as a bucket of crabs can be.
It's only when deliberate uncensor-ings are made that some form of usefulness can be clawed back.
https://huggingface.co/ehartford
https://huggingface.co/ehartford/Wizard-Vicuna-13B-Uncensore...
I wish the overwhelming tide of reaction had been to uncover what it does well and how to make it better. Sure, look for holes and bugs, but in a serious and constructive way.
Anybody looking for ways to critique AI found copious clickbaity examples up and down this forum.
Go in your history.
Pick a conversation.
Copy and paste the first prompt and put it into ChatGPT and compare the results.
I think what may be happening is that your expectations have shifted. When a tool proves useful we're used to being able to use it reliably, but with knowledge exfiltration it is different because our queries keep changing. It's not the same thing as using sandpaper on a piece of oak. Overall, I've been very impressed with ChatGPT and continue to use it for all sorts of things.
That said, tricking it into giving you the answer on political things is an art that requires patience. It is too, uh, cautious.
1 - https://arxiv.org/pdf/2203.02155.pdf
2 - https://crfm.stanford.edu/helm/v0.1.0/?group=question_answer...
Edit: I get you're trying to make a point by posting a stupid question, but this isn't a good way to make a point.
Just say what you want to say. You don't need to hide it behind sarcasm.
I wonder why you can’t get it to give approximations. I have not encountered difficulties there. I asked it to approximate the weight of argon in the atmosphere and it performed a fairly standard back of the envelope calculation. Do you want to share a particular prompt, so I can try and reproduce the issue you're experiencing?
But now OpenAI is lobbying the government to make it unlawful for others to do their own thing and that changes the equation significantly.