Skip to main content
Ryan Orban

Ryan Orban

Subject
1 entry

Dialogue Systems

Bookmarks

  1. Building Safer Dialogue Agents (DeepMind / Sparrow)

    DeepMind's blog post on Sparrow — a dialogue agent trained with reinforcement learning from human feedback and rules to be helpful, harmless, and honest. An early published account of RLHF-based safety fine-tuning for conversational AI.

All bookmarks