Skip to main content
Ryan Orban

Ryan Orban

Subject
1 entry

Llm Safety

Bookmarks

  1. Machine Generated Text: A Comprehensive Survey of Threat Models and Detection Methods

    Comprehensive survey by Crothers, Japkowicz, and Viktor mapping the threat models and detection methods for machine-generated text, with emphasis on fairness and robustness. As generative models like ChatGPT become widely accessible, the gap between reliable detection and adversarial evasion has become one of the central unsolved problems in AI safety and content integrity.

All bookmarks