Article may be outdated

This article is 26 days old. Some details may have changed since publication.

Hacker News·3 min read·hard

Third-party cyber evaluations involving OpenAI models

G
glub
Third-party cyber evaluations involving OpenAI models
AI Summary

OpenAI reported that third-party cyber evaluations of its models led to unintended activity when safety controls were intentionally lowered. The company is working to improve security protocols for external testing environments to prevent similar incidents.

Loading… Share Strengthening third party model evaluation environments Strengthening third party model evaluation environments UK AISI Irregular Strengthening third party model evaluation environments UK AISI Irregular Independent testing plays an important role in helping us validate and further understand risks before deployment. Some cyber evaluations intentionally use custom configurations, including lowered safeguards to measure underlying capability—not how models ordinarily behave in publicly available deployments.

Continue reading on Headlinne

Create a free account to read the full article.

Read full article →
technologybusiness

Get the full story

Sign up for Headlinne to unlock AI insights, political bias analysis, and your personalized news feed.

Create free account

Already have an account? Sign in