Building smarter maps with GPT-4o vision fine-tuning
Building smarter maps with GPT-4o vision fine-tuning
Developers can now fine-tune GPT-4o with images and text to improve vision capabilities
Building smarter maps with GPT-4o vision fine-tuning
Our latest image generation model is now available in the API via ‘gpt-image-1’—enabling developers and businesses to build professional-grade, customizable visuals directly into their own tools and platforms.
At OpenAI, we have long believed image generation should be a primary capability of our language models. That’s why we’ve built our most advanced image generator yet into GPT‑4o.
Be My Eyes uses GPT-4 to transform visual accessibility.
We’ve created GPT-4, the latest milestone in OpenAI’s effort in scaling up deep learning. GPT-4 is a large multimodal model (accepting image and text inputs, emitting text outputs) that, while less capable than humans in many real-world scenarios, exhibits human-level performance on various professional and academic benchmarks.
Introducing GPT-4. 1 in the API—a new family of models with across-the-board improvements, including major gains in coding, instruction following, and long-context understanding.
We’re announcing GPT-4 Omni, our new flagship model which can reason across audio, vision, and text in real time.
GPT-4 Turbo with 128K context and lower prices, the new Assistants API, GPT-4 Turbo with Vision, DALL·E 3 API, and more.
Introducing GPT-5 in our API platform—offering high reasoning performance, new controls for devs, and best-in-class results on real coding tasks.
This 2025 launch introduced a faster ChatGPT Images experience and GPT‑Image‑1.5 in the API. Explore the latest ChatGPT Images 2.5.
The new ChatGPT Images is powered by our flagship image generation model, delivering more precise edits, consistent details, and image generation up to 4× faster. The upgraded model is rolling out to all ChatGPT users today and is also available in the API as GPT-Image-1.