Is GPT-4.1 nano a dated version of GPT-4.1?
No. nano, mini, and GPT-4.1 are different models in the same family, with nano focused on low-latency small tasks. When calling it, use gpt-4.1-nano; do not directly apply other models' coding scores, vision performance, or maximum output numbers to it.
With a million-token context, can it directly perform complex contract reviews?
It can perform targeted searches within the provided contract text, but complex reviews also involve clause relationships, exceptions, and cross-document reasoning. nano is better suited to extracting specified fields or locating paragraphs; when comprehensive risk assessment is needed, choose a more capable model and retain human review.
How can I make nano look at images instead of answering only text questions?
In Chat Completions messages, write the user content as an array of content blocks, including both a text question and an image_url image address. The output is a text response based on the image; if you want to generate or edit images, choose a dedicated image model.
How should I choose between Responses and Chat Completions?
Both can specify gpt-4.1-nano. Use Chat Completions if you already have a messages conversation structure; use Responses input if you use response objects and event-stream processing. Do not mix the content structures of the two entry points, and the client must read results according to the corresponding response format.
Must I save the complete history myself for multi-turn conversations?
Standard conversation calls are usually organized by the application through historical messages; if you want to simplify session maintenance, you can choose a managed conversation entry point, set stateful, and include the returned id in subsequent requests. Key rules should still be expressed clearly, and long conversations cannot guarantee that all early details are always retained accurately.