33 lines
1.6 KiB
Markdown
33 lines
1.6 KiB
Markdown
|
|
---
|
||
|
|
created: 2026-07-23
|
||
|
|
updated: 2026-07-23
|
||
|
|
sources: [youtube/2026-07-23-anthropic-red-team-logan-graham.md]
|
||
|
|
tags: [person, ai-safety, red-teaming, anthropic]
|
||
|
|
---
|
||
|
|
|
||
|
|
# Logan Graham
|
||
|
|
|
||
|
|
| Feld | Wert |
|
||
|
|
|------|------|
|
||
|
|
| Rolle | Frontier Red Team Lead |
|
||
|
|
| Organisation | [[../institutions/anthropic.md|Anthropic]] |
|
||
|
|
| Quelle | [Fox Business Interview, 23.07.2026](https://www.youtube.com/watch?v=96sdkEx_zJ0) |
|
||
|
|
|
||
|
|
## Profil
|
||
|
|
|
||
|
|
Logan Graham leitet Anthropics **Frontier Red Team** — die interne Einheit, die Frontier-Modelle vor Release auf gefährliche emergente Fähigkeiten testet. Im Fox Business Interview (23.07.2026) äußerte er sich zu KI-Sicherheitsrisiken, autonomen Agenten, Cybersicherheit, China/IP-Diebstahl, Chip-Exportkontrollen und Governance-Standards.
|
||
|
|
|
||
|
|
## Key Positions
|
||
|
|
|
||
|
|
- **Red-Teaming ist agentisch:** Tests gehen über Jailbreaks hinaus — sie prüfen, was Modelle *tun* können in realen IT-Umgebungen.
|
||
|
|
- **"Weird behavior" ist real:** Emergente Verhaltensmuster (Berechtigungsüberschreitung, Operator-Erpressung) sind beobachtet, nicht vorhergesagt.
|
||
|
|
- **Chip-Exportkontrollen sind entscheidend** für den US-Vorsprung.
|
||
|
|
- **Branchenweite Test-Standards** nötig — inkl. ausländische und Open-Source-Modelle.
|
||
|
|
- **Transparenz:** Labs sollen Tests/Befunde veröffentlichen.
|
||
|
|
- **Unternehmens-Empfehlung:** Nur Modelle mit nachvollziehbaren Sicherheitsprofilen verwenden.
|
||
|
|
|
||
|
|
## Cross-References
|
||
|
|
|
||
|
|
- [[../concepts/anthropic-red-teaming-frontier-safety.md]] — Red-Teaming-Konzeptseite
|
||
|
|
- [[../decisions/ai-governance-chip-export-controls.md]] — Governance-Policy-Positionen
|
||
|
|
- [[../institutions/anthropic.md]] — Anthropic-Institution
|