shot-button
Subscription Subscription
Home > Technology News > What exactly is ChatGPT All you need to know about the artificial intelligence chatbot

What exactly is ChatGPT? All you need to know about the artificial intelligence chatbot

Updated on: 15 January,2023 04:54 PM IST  |  New Delhi
IANS |

The dialogue format makes it possible for ChatGPT to answer follow-up questions, admit its mistakes, challenge incorrect premises, and reject inappropriate requests

What exactly is ChatGPT? All you need to know about the artificial intelligence chatbot

Image for representational purpose only. Photo Courtesy: istock

As artificial intelligence takes over many aspects of our life, the most recent technology to take over the internet is of AI-driven chatbot called ChatGPT that writes poems and essays and makes humourous comments through it. The conversational AI medium has reportedly opened several avenues, but has to be handled with care. 

According to OpenAI, the company behind chatGPT, they have trained an AI model which interacts in a conversational way.


The dialogue format makes it possible for ChatGPT to answer follow-up questions, admit its mistakes, challenge incorrect premises, and reject inappropriate requests.


ChatGPT is a sibling model to "InstructGPT", which is trained to follow an instruction in a prompt and provide a detailed response, according to OpenAI which was acquired by Microsoft for $1 billion.


This is how it works.

The company trained the model using 'Reinforcement Learning from Human Feedback' (RLHF), using the same methods as InstructGPT, but with slight differences in the data collection setup.

"We trained an initial model using supervised fine-tuning: human AI trainers provided conversations in which they played both sides - the user and an AI assistant," says OpenAI.

The teams gave the trainers access to model-written suggestions to help them compose their responses.

"We mixed this new dialogue dataset with the InstructGPT dataset, which we transformed into a dialogue format," the company informed.

To create a reward model for reinforcement learning, it took conversations that AI trainers had with the chatbot.

"We randomly selected a model-written message, sampled several alternative completions, and had AI trainers rank them. Using these reward models, we can fine-tune the model using 'Proximal Policy Optimisation'. We performed several iterations of this process," explained OpenAI.

What are its limitations?

ChatGPT sometimes writes plausible-sounding but incorrect or nonsensical answers.

According to the company, fixing this issue is challenging, as during RL training, there's currently no source of truth and training the model to be more cautious causes it to decline questions that it can answer correctly.

Also, supervised training misleads the model because the "ideal answer depends on what the model knows, rather than what the human demonstrator knows".

ChatGPT is sensitive to tweaks to the input phrasing or attempting the same prompt multiple times. For example, given one phrasing of a question, the model can claim to not know the answer, but given a slight rephrase, can answer correctly, according to OpenAI.

The model is often excessively verbose and overuses certain phrases, such as restating that it's a language model trained by OpenAI.

"These issues arise from biases in the training data (trainers prefer longer answers that look more comprehensive) and well-known over-optimisation issues," the company admitted.

"While we've made efforts to make the model refuse inappropriate requests, it will sometimes respond to harmful instructions or exhibit biased behaviour. We're using the Moderation API to warn or block certain types of unsafe content, but we expect it to have some false negatives and positives for now," it added.

The company is currently collecting user feedback.

Also Read: NFT nuptials: How this Pune couple got married on the blockchain with 'digital rings' and a 'digital priest'

This story has been sourced from a third party syndicated feed, agencies. Mid-day accepts no responsibility or liability for its dependability, trustworthiness, reliability and data of the text. Mid-day management/mid-day.com reserves the sole right to alter, delete or remove (without notice) the content in its absolute discretion for any reason whatsoever

"Exciting news! Mid-day is now on WhatsApp Channels Subscribe today by clicking the link and stay updated with the latest news!" Click here!

Do you frequently use AI art and writing softwares?

Register for FREE
to continue reading !

This is not a paywall.
However, your registration helps us understand your preferences better and enables us to provide insightful and credible journalism for all our readers.

Mid-Day Web Stories

Mid-Day Web Stories

This website uses cookie or similar technologies, to enhance your browsing experience and provide personalised recommendations. By continuing to use our website, you agree to our Privacy Policy and Cookie Policy. OK