Advancing the price-performance frontier with GPT-5.6
OpenAI releases GPT-5.6 with improved price-performance ratios for the Luna and Terra models. These efficiencies are designed to help enterprises scale AI workflows more affordably.
Le meilleur de l'écosystème IA et MCP, sélectionné chaque jour.
Sources
OpenAI releases GPT-5.6 with improved price-performance ratios for the Luna and Terra models. These efficiencies are designed to help enterprises scale AI workflows more affordably.
OpenAI demonstrates how specific API settings for reasoning and compaction tripled GPT-5.6 performance on the ARC-AGI-3 benchmark. This highlights a critical optimization path for improving model efficiency and reasoning capabilities in complex problem-solving tasks.
OpenAI introduces GPT-5.6, focusing on maximizing the ratio of intelligence to compute. The update optimizes inference and agentic workflows to deliver higher utility per dollar for developers.
OpenAI explores how AI coding agents are modernizing scientific computing. The report highlights significant accelerations in software development and discovery within the genomics field.
OpenAI launches Presence, an enterprise AI agent platform designed to deploy trusted voice and chat agents for internal and customer-facing workflows.
OpenAI and Hugging Face have released findings on a security incident encountered during model evaluation. The report details advanced cyber capabilities and provides defensive lessons for the AI research community.
OpenAI discusses the unique safety risks and failure modes associated with long-horizon models. The post outlines iterative deployment strategies and new safeguards necessary for AI that operates over extended timeframes.
OpenAI's creative team is leveraging Codex to build custom tools and accelerate prototyping. This demonstrates the practical application of LLMs in enhancing design and ideation workflows within a professional creative environment.
OpenAI introduces GPT-Red, an automated red teaming system using self-play to enhance AI safety and robustness. The system specifically targets improvements in alignment and resistance to prompt injection attacks.
GPT-5.6 is now the default engine for Microsoft 365 Copilot, enhancing capabilities across the productivity suite. This update brings higher quality output and faster performance to Word, Excel, and PowerPoint.
OpenAI introduces ChatGPT Work, an agent capable of executing long-running tasks across various applications and files. This represents a shift toward more autonomous, project-based AI agents that can manage goals over several hours.
OpenAI has released GPT-5.6, offering higher intelligence per token and improved performance efficiency. This update aims to provide more on-demand capability for complex, high-ambition workloads.
OpenAI analyzes reliability issues in the SWE-Bench Pro coding benchmark. The findings highlight the need for more accurate evaluation methods for AI coding models.
OpenAI has launched GPT-Live, a new generation of voice models designed for more natural, real-time human-AI interaction. The technology is now integrated into ChatGPT Voice, significantly reducing latency and improving conversational fluidity.
OpenAI introduces GeneBench-Pro, a new benchmark for AI performance in genomics and biology. It uses complex real-world datasets to evaluate scientific research capabilities.
OpenAI engineers leveraged large-scale core dump analysis to resolve rare infrastructure crashes. The process revealed a combination of hardware faults and a legacy software bug that had persisted for 18 years.
OpenAI has previewed GPT-5.6 Sol, featuring significant improvements in coding, science, and cybersecurity capabilities. The model is launched alongside a new, advanced safety stack to mitigate high-capability risks.
OpenAI releases a research paper detailing how AI agents are managing complex, long-term tasks and increasing productivity across professional roles. It highlights the shift from simple chat interactions to autonomous agentic workflows.
OpenAI and Broadcom have introduced 'Jalapeño', a custom AI chip specifically engineered for LLM inference. This hardware optimization aims to significantly increase performance, energy efficiency, and scalability for large-scale AI deployments.
GPT-5 Pro demonstrates advanced reasoning capabilities by solving a long-standing immunology mystery regarding T cell behavior. The result highlights the model's utility in specialized scientific research.