İçeriğe atla
Hikayeler
Hikaye
Gelişiyor

OpenAI, Astra Modelinde Kritik Siber Güvenlik Riski Tespit Etti, Geliştirmeyi Durdurdu

Özet · AI üretimi

OpenAI, Cuma günü yaptığı açıklamada, geliştirme aşamasındaki yapay zekâ modeli Astra’nın 'kritik' siber güvenlik yeteneklerine sahip olma ihtimalini dışlayamadığını bildirdi. Bu, modelin sıfır gün açıkları olarak bilinen ciddi yazılım zafiyetlerini otonom bir şekilde tespit edip istismar edebileceği anlamına geliyor. Şirket, iç güvenlik kılavuzunda tanımlanan 'kritik' eşiğin aşılması nedeniyle Astra’nın bazı dahili geliştirme süreçlerini durdurdu ve güvenlik protokollerini devreye aldı. OpenAI’ın kılavuzuna göre, bir modelin bu eşiğe ulaşması, kötü niyetli kullanım potansiyelinin çok yüksek olduğunu gösteriyor. Astra’daki yetenek, otonom siber saldırılar düzenlenmesi gibi riskler barındırıyor. Geliştirmenin askıya alınması, şirketin bu tür riskleri ciddiye aldığını yansıtıyor. Bu gelişme, ileri yapay zekâ modellerinin siber güvenlik alanındaki çift yönlü kullanım risklerini vurguluyor. Kritik eşiğe ulaşan bir modelin kontrolsüz kalması, küresel siber güvenlik dengesini tehdit edebilir. OpenAI’ın proaktif adımı, sektörde güvenlik odaklı geliştirme yaklaşımının önemini ortaya koyuyor.

Başlangıç 08 Ağu 14:32 1 olay Güncellendi 3 sa önce
Paylaş
Bağlam · AI üretimi

Bağlam, hikayenin etrafındaki ülke + lider + komşu hikaye ağına dayanılarak AI tarafından üretildi. Olgu içerikleri için her zaman üstteki kaynak linklerine başvurun.

Bu gündemi takip et

gelişmelerini kaçırma — ücretsiz kaydol, günlük brifinginde gör.

Bu gündeme tepki ver:

Zaman çizelgesi

en güncel: 3 sa önce
  1. Güvenlik08 Ağu 14:32

    OpenAI flags possible critical cybersecurity risk in upcoming model Astra, tightens controls

    OpenAI said on Friday it cannot rule out that its upcoming AI model, Astra, has “critical” cybersecurity capabilities, prompting the startup to pause some internal development and trigger safety protocols. Under OpenAI’s safety guidelines, a model reaches the “critical” threshold if it can autonomously identify and exploit severe, real-world software vulnerabilities, known as zero-day exploits, or execute complex cyberattacks against highly secure targets without human intervention. Here are some details on Astra: This follows an exclusive report by Reuters that OpenAI has discovered more instances in which autonomous agents have escaped containment as the company expands its investigation of the hacking incident at tech firm Hugging Face that drew global attention in July. In the last few weeks, OpenAI, Anthropic and Meta Platforms have disclosed that their AI models broke into other companies’ systems during cybersecurity testing, highlighting how advancing AI capabilities are straining developers’ ability to keep their systems contained. Preliminary evaluations over the past several days, along with outside expert assessments, indicated Astra may be capable of performing increasingly sophisticated cyber tasks autonomously, OpenAI said. “While we continue to benchmark and assess this model, our preliminary evaluations indicate strong enough performance that we cannot rule out ‘critical’ capability level at this time,” the ChatGPT maker said. In response to the preliminary findings, OpenAI said it has scaled up security controls and paused internal activities involving Astra that do not meet its newly strengthened security requirements. Astra’s development will be moved into isolated testing environments with restricted network access and sandboxed execution. CEO Sam Altman said on X that OpenAI is working to make Astra generally available, as the company does “not think it is a good strategy to keep powerful models to a chosen few”. OpenAI also clarified that Astra was not involved in the hack targeting the AI platform Hugging Face. It will partner with government agencies and select AI safety organisations to test the model’s capabilities.