Prompt Separation
Prompt-Separation Voice-typed podcast prompt transcripts decomposed into structured fields: discrete prompts (asks), a list of context chunks, and free-form host_notes. The dataset supports training a small model that, given a single voice-typed message, recovers the structured fields an AI host would consume — separating "what is the user actually asking?" from "what is the surrounding context?" from "how should the response be shaped?". Source Prompts come from the… See the full description on the dataset page: https://huggingface.co/datasets/danielrosehill/Prompt-Separation.
View on Hugging Face