Thinking with images
OpenAI o3 and o4-mini models achieve visual reasoning breakthrough by incorporating images in chain-of-thought processing.
OpenAI o3 and o4-mini models achieve visual reasoning breakthrough by incorporating images in chain-of-thought processing.
Video generation in Gemini and Whisk using Veo 2. Feature announcement for generative models.
OpenAI releases GPT-4.1 API with improvements in coding, instruction following, and long-context understanding, including first nano model.
BrowseComp is a benchmark for evaluating and comparing browsing agents' ability to navigate and retrieve information from web environments.
PaperBench is a benchmark for evaluating AI agents' ability to replicate state-of-the-art AI research end-to-end.
Zendesk pilots new AI agents powered by OpenAI models that autonomously plan and execute customer service responses beyond intent-based conversation management.
OpenAI shares updates on Cybersecurity Grant Program, bug bounties, and security initiatives including research on prompt injection and AGI safety.
Hebbia's Matrix platform uses multiple AI agents in parallel to automate complex financial and legal workflows, orchestrating OpenAI's o3-mini and o1 models.
OpenAI Academy expands with free online resource hub for AI literacy, best practices, and tools for diverse audiences.
OpenAI and MIT Media Lab research collaboration studying affective use and emotional well-being impacts of ChatGPT interactions.
Booking.com integrates OpenAI LLMs to deliver smarter search, faster support, and intent-driven travel experiences at scale.
ChatGPT for Business March 2025 updates focus on interactive, customizable, and agentic capabilities for teams.
Native image generation in Gemini 2.0 Flash for developers. Feature addition for developer use.
OpenAI releases first building blocks for developers and enterprises to build useful and reliable AI agents with complex multi-step task capabilities.
Nubank leverages OpenAI solutions to enhance customer experience and internal efficiency for 114M users across Latin America.
Factory Platform uses OpenAI o1, o3-mini, and GPT-4o for software development workflows, achieving 20% engineering cycle acceleration.
LaunchDarkly CPTO discusses AI-powered product management and the evolving role of product managers in AI applications.
Mercari marketplace uses GPT-4o mini and GPT-4 for product listing enhancement and seller support features.
OpenAI GPT-4.5 system card detailing risk assessment, preparedness scorecard, and deployment criteria for the new model.
OpenAI releases GPT-4.5, largest model yet with improved pattern recognition and natural interaction, available to Pro users and developers.
OpenAI Deep Research System Card detailing safety evaluations, red teaming, and risk mitigations before release.
Estonia and OpenAI partnership to provide ChatGPT Edu access to secondary school students and teachers.
Survey of ChatGPT adoption among US college students and implications for workforce readiness.
SWE-Lancer benchmark evaluates frontier LLMs on real-world freelance software engineering tasks and earnings potential.
General overview of large language model architectures and characteristics similar to ChatGPT.
Safety evaluation report for OpenAI o3-mini model including red teaming and preparedness framework assessments.
ChatGPT Government edition for U.S. government agencies to access OpenAI models.
OpenAI safety documentation for AI systems covering prompt injection, jailbreak, privacy, and red teaming mitigations.
Universal interface enabling AI agents to interact with digital environments and applications.
Research on improving adversarial robustness by trading inference-time compute in language models.
New alignment strategy for o1 models using reasoning to teach safety specifications.
Benchmark measuring LLM factuality and hallucination reduction. Evaluation tool for LLM reliability.
Gemini 2.0 multimodal model positioned for agentic applications. Core LLM release for agent-centric era.
Zalando uses GPT-4o mini to power customer service assistant improving retail experience.
OpenAI's Sora video generation model technical overview. Generates video from text/image/video inputs up to 1080p.
OpenAI o1 safety report detailing red teaming, risk evaluations, and preparedness framework.
Red teaming methodology combining human and AI-based safety testing for model evaluation.
Fine-tuning GPT-4o vision for improved map building and spatial understanding.
ChatGPT search feature provides web-sourced answers with citations.
Case study: Promega accelerates manufacturing and sales operations using ChatGPT.
Simplified continuous-time consistency models achieving comparable quality to diffusion models with only two sampling steps.
OpenAI o1 reasoning models applied to coding, strategy, and research problems. Video overview of capabilities.
Research analyzing ChatGPT response fairness based on user names using privacy-preserving AI assistants.
MLE-bench: Benchmark for evaluating AI agents on machine learning engineering tasks.
OpenAI announces fine-tuning API now supports vision models (GPT-4o) for improved image and text capabilities.
OpenAI API feature providing automatic cost discounts for repeated input tokens through prompt caching.
Altera uses GPT-4o to enable human-AI collaboration in their platform.
OpenAI releases multimodal moderation model based on GPT-4o for detecting harmful text and images in developer applications.
Google releases updated production-ready Gemini models with reduced pricing and increased rate limits.
Mercado Libre introduces Verdi, an AI developer platform powered by GPT-4o.