Multi-Modal GPT-4: The Good, the Bad, & the Ugly

OpenAI is about to launch multi-modal GPT-4 with expanded capabilities that were unimaginable even this time last year. Unfortunately, as with all new technology there are both amazing possibilities and enormous risks.

OpenAI outlines these risks in a recent paper covering the early testing of their image to text generator: GPT-4V(ision). I break down the paper in the video, but the TL:DW version is:

• This model is prone to hallucinations, inaccuracies, and biases like all other GenAI models

• The bot consistently provided stereotyped answers, responded with ungrounded inferences, and inconsistently answered scientific and medical prompts

• Compounding these issues - the new capabilities make it easier to create disinformation campaigns through the pairing of text to image, known to make disinformation seem more trustworthy

• Mitigation strategies have been put in place - like the bot blocking certain types of prompts - though it was possible to sidestep these strategies with user effort

• There was no mention of an updated user experience for ChatGPT-4 to make it easier to identify these potentially harmful and damaging outputs

• The risks outlined in the paper underscores the need for building AI literacy and skills for critical evaluation of these tools and their limitations/biases

Right now it is imperative for us to take a balanced approach to the adoption and application of GenAI to limit the foreseen and unforeseen consequences.
Previous
Previous

Thanksgiving Activities Prompts

Next
Next

Higher Ed Workshop - Two Things