Prompt Refusal

FromData Skeptic

Start listening View podcast show

Prompt Refusal

FromData Skeptic

ratings:

Length:

44 minutes

Released:

Jul 24, 2023

Format:

Podcast episode

Description

The creators of large language models impose restrictions on some of the types of requests one might make of them. LLMs commonly refuse to give advice on committing crimes, producting adult content, or respond with any details about a variety of sensitive subjects. As with any content filtering system, you have false positives and false negatives. Today's interview with Max Reuter and William Schulze discusses their paper "I'm Afraid I Can't Do That: Predicting Prompt Refusal in Black-Box Generative Language Models". In this work, they explore what types of prompts get refused and build a machine learning classifier adept at predicting if a particular prompt will be refused or not.

Released:

Jul 24, 2023

Format:

Podcast episode

Titles in the series (100)

Data Skeptic is a data science podcast exploring machine learning, statistics, artificial intelligence, and other data topics through short tutorials and interviews with domain experts.

Skip carousel

More Episodes from Data Skeptic

Skip carousel

Related podcast episodes

Skip carousel

Discover this podcast and so much more

Prompt Refusal

Prompt Refusal

Description

Titles in the series (100)

More Episodes from Data Skeptic

Related podcast episodes