Tag
139 articles
This explainer explores the shift from traditional software development to AI development in the era of generative and agentic AI, examining key skills required and identifying gaps in current frameworks.
A new analysis examines four falsifiable conditions that must be met for agentic coding to replace junior engineers, based on evidence from METR, OpenAI, DORA, and Stanford.
OpenAI's CFO Sarah Friar outlines how advances across chips, compute, models, and products compound to deliver more useful intelligence at greater scale and lower cost.
Nvidia's research demonstrates that AI agents can perform well through fine-tuning and harness design, rather than relying solely on advanced models.
This article explains the concept of AI model factories and their significance in modern AI development, using Nvidia's acquisition of Poolside as a case study.
OpenAI has slowed its AI development pace amid increasing competition and regulatory pressure, pausing some activities to strengthen security and safeguards.
OpenAI is implementing enhanced monitoring, alignment, and security measures for its frontier AI models, slowing development pace to prioritize safety and responsible deployment.
OpenAI has implemented new security safeguards following the Hugging Face breach, emphasizing enhanced monitoring and alignment protocols during AI development.
OpenAI has overhauled its safety protocols after its Astra AI model demonstrated unexpected cyber capabilities, prompting a pause in training runs and enhanced internal safeguards.
Warp introduces Warp Factories, a new infrastructure system designed to simplify AI software development by providing pre-configured environments and automated workflows.
This article explains the concept of catastrophic AI risks and why AI safety is crucial. It explores the recent decision by OpenAI to dissolve a team focused on identifying such risks, and why this matters for the future of AI development.
Tim O'Reilly, founder of O'Reilly Media, criticizes major AI labs for focusing on technical benchmarks rather than user needs, advocating instead for open-source AI development.