gpt-4.1
OpenAI model with major improvements in coding and instruction following. Completes 54.6% of SWE-bench Verified tasks vs 33.2% for GPT-4o. Features 1M token context window and improved long-context comprehension.