![]() OpenAI, bilgisayar kullanımında insanüstü performansa ulaştıOpenAI achieves superhuman performance in computer useHabere gitRead the article Türkçe (otomatik çeviri) English (automatic translation) |
ÖzetSummaryLatent Space podcast'ında OpenAI'nin DevDay 2026 etkinliği ele alınıyor. CUA ekibinin lideri Ari Weinstein, Computer Use'un '180 derece farklı' bir noktaya ulaştığını ve bazı görevlerde ortalama insanları geride bırakabildiğini savunuyor. API ürün lideri Nikunj Handa ise bir haftada piyasaya sürülen Decisions API'sini (Jev rakibi) ve asenkron araç çağrısı, UltraFast çıkarım ile istem önbellekleme gibi geniş kapsamlı ajan yığını yenilemelerini anlatıyor.The Latent Space podcast covers OpenAI's DevDay 2026 event. Ari Weinstein, leader of the CUA team, argues that Computer Use has reached a '180-degree different' point and can outperform average humans in certain tasks. API product leader Nikunj Handa discusses the Decisions API (a Jev competitor) launched within a week, along with broad agent stack updates including asynchronous tool calling, UltraFast inference, and prompt caching. |
Neden ÖnemliWhy it mattersComputer Use, tam dijital iş gücüne en yakın mevcut yapay zeka yeteneği. OpenAI'nin süperinsan görev tamamlama iddiası, rekabet çerçevesini 'yapabiliyor mu?' sorusundan 'ne kadar hızlı ve ucuz yapıyor?' sorusuna kaydırıyor. Bir haftada tamamlanan Decisions API süreci, öncü laboratuvarların farklılaşmasının artık ham model atılımlarından çok paketleme ve gecikme süresine bağlı olduğunu gösteriyor. Bu durum, bağımsız ajan girişimlerinin rekabet avantajını daraltıyor.Computer Use is the closest current AI capability to a fully digital workforce. OpenAI's claim of superhuman task completion shifts the competitive framework from 'can it do it?' to 'how fast and cheaply can it do it?'. The fact that the Decisions API was completed in a week shows that leading labs' differentiation now relies more on packaging and latency than on raw model breakthroughs. This narrows the competitive advantage of independent agent startups. |
Öne ÇıkanlarHighlights
|
EleştiriCritical takeTüm içerik, konukların OpenAI'nin kendi ekibinden oluştuğu bir öz tanıtım podcast'i. Bu nedenle '180 derece farklı' ve 'süperinsan' gibi ifadeler, bağımsız bir kıyaslama içermeyen doğrulanmamış pazarlama dili. 'Süperinsan' çerçevesi 'bazı görevlerde ortalama insandan hızlı' şeklinde yumuşatılmış. Bu, başlığın ima ettiğinden çok daha zayıf ve az spesifik bir iddia.The entire content is a self-promotional podcast featuring guests from OpenAI's own team. Therefore, phrases like '180-degree different' and 'superhuman' are unverified marketing language lacking independent benchmarking. The 'superhuman' framing has been softened to 'faster than the average human in some tasks.' This is a much weaker and less specific claim than the title implies. |
