Global ETD Search

Search theses and dissertations gathered from participating repositories worldwide. Every result links back to the library that holds it. No account is needed.

Results

Showing 1 to 1 of 1 for “"llm-human collaboration"”.

  1. Adversarial Risks and Stereotype Mitigation at Scale in Generative Models

    … rely on a `natural channel' of code, such as human-readable tokens and structure, that adversaries can exploit with minimal perturbations. These attacks expose the fragility of state-of-the-art PL models, highlighting how superficial patterns and hidden assumptions in training data can lead to …

    vt Repository record for Adversarial Risks and Stereotype Mitigation at Scale in Generative Models (opens in a new tab)