Skip to content
Tag

Transparency

5 stories

2 min read

OpenAI confirms the German wiki incident

The company treats it as a misalignment case and separates it from the Hugging Face breach. But the confirmation was not volunteered: executives knew weeks earlier and the statement followed the Reuters story.

Safety
8 min read

How AI text detection works, and why it is hard

Detection tools produce probability, not certainty. Four families of methods, the three limits of watermarking, the cost of a false positive and the extra difficulty Turkish adds — a practical framework for how an organisation should use these tools.

Safety
3 min read

The secret model-review framework goes to court

Protect Democracy has sued four federal agencies to force the release of the framework used to review frontier models before launch. The framework is not classified, yet it still is not being shared.

Policy