Is GPT-4.1 nano a dated version of GPT-4.1?
No. nano, mini, and GPT-4.1 are different models in the same family, with nano focused on low-latency small tasks. When calling it, use gpt-4.1-nano; do not directly apply other models' coding scores, vision performance, or maximum output numbers to it.
With a million-token context, can it directly handle complex contract review?
It can perform targeted searches within the provided contract text, but complex review also involves clause relationships, exceptions, and cross-document reasoning. nano is better suited to extracting specified fields or locating paragraphs; when a complete risk assessment is needed, choose a more capable model and retain human review.
How can I make nano look at images instead of only answering text questions?
In Chat Completions messages, write the user content as an array of content blocks, including both the text question and the image_url image address. The output is a text answer based on the image; if you want to generate or edit images, choose a dedicated image model.
How should I choose between Responses and Chat Completions?
Both can specify gpt-4.1-nano. Use Chat Completions if you already have a messages conversation structure; use Responses input if you use response objects and event-stream processing. Do not mix the content structures of the two entry points, and the client must read results according to the corresponding response format.
Must I save the full history myself for multi-turn conversations?
When using Chat Completions, place relevant history in messages; when using Responses, organize input and related conversation content according to the documentation. Provide the latest materials, revision goals, and key constraints in each turn; for longer tasks, retain phased summaries and a final version that can be checked independently.