Can GPT-4o mini view images and generate images?
It can generate text answers by combining images and text prompts, making it suitable for image descriptions and visible information extraction, but it is not an image generation model. In Chat Completions, you can combine text and image_url blocks in the content array; if you need an actual finished image, use a dedicated image generation or editing model.
How do I choose between Responses and Chat Completions?
Applications that already use a messages conversation structure can use /openai/chat/completions and read responses from choices; responsive workflows can use /openai/responses and submit input. Both use model: gpt-4o-mini, but their input organization and response parsing differ.
Do I need to resend the history for every turn of a multi-turn conversation?
When using Chat Completions directly, you should include the required history in messages. If you want to simplify session management, you can use the AI Chat session interface, enable stateful, and include the returned id in subsequent requests. Session persistence is an interface feature and does not mean the model has unlimited context.
Will function calling directly execute my business code?
Defining a function alone will not automatically execute code. The model can return a function name and parameters; the application is responsible for validating parameters, performing the corresponding operation, and returning the result. Tasks such as querying orders and sending emails should each have their own permissions and execution logic configured, especially do not omit confirmation steps for write operations.
Can I directly send a PDF as an image?
PDF files and image inputs are not the same type of request. For GPT-4o mini, clear images and extracted text are the primary input methods described here. When processing PDFs, you can first extract the body text or convert relevant pages to images, then submit the content according to goals such as summarization, field extraction, or question answering.