293 models ranked for CI/CD and deployment automation. Scored with bonuses for function calling (pipeline triggers), JSON mode (config files), reasoning (debugging builds), large context, streaming, and web search.
| # | Model | Score |
|---|---|---|
| 1 | GPT-5.4 ProOpenAI | 91 |
| 2 | GPT-5.2 ProOpenAI | 90 |
| 3 | GPT-5 ProOpenAI | 90 |
| 4 | o3 ProOpenAI | 82 |
| 5 | Claude Opus 4.1Anthropic | 81 |
| 6 | o3 Deep ResearchOpenAI | 74 |
| 7 | Claude Opus 4.6Anthropic | 71 |
| 8 | Claude Opus 4Anthropic | 76 |
| 9 | Claude Opus 4.5Anthropic | 70 |
| 10 | GPT-5.4OpenAI | 70 |
| 11 | Claude Sonnet 4.5Anthropic | 69 |
| 12 | Qwen3 VL 30B A3B ThinkingAlibaba | 69 |
| 13 | Qwen3 VL 235B A22B ThinkingAlibaba | 69 |
| 14 | GPT-5.2OpenAI | 68 |
| 15 | Claude Sonnet 4.6Anthropic | 68 |
| 16 | GPT-5.1OpenAI | 67 |
| 17 | o1-proOpenAI | 77 |
| 18 | GPT-5.3-CodexOpenAI | 67 |
| 19 | GPT-5.2-CodexOpenAI | 67 |
| 20 | GPT-5OpenAI | 67 |
| 21 | Gemini 3.1 Pro Preview Custom ToolsGoogle | 68 |
| 22 | Gemini 3.1 Pro PreviewGoogle | 68 |
| 23 | Gemini 3 Pro PreviewGoogle | 68 |
| 24 | o4 Mini Deep ResearchOpenAI | 66 |
| 25 | GPT-5.1-Codex-MaxOpenAI | 66 |
| 26 | GPT-5 MiniOpenAI | 65 |
| 27 | GPT-5 NanoOpenAI | 64 |
| 28 | Gemini 3 Flash PreviewGoogle | 66 |
| 29 | Grok 4.1 FastxAI | 64 |
| 30 | Grok 4 FastxAI | 64 |
Generate GitHub Actions, GitLab CI, Jenkins, and CircleCI pipeline configurations. JSON mode produces valid YAML-compatible structured output.
Analyze build logs, identify slow steps, and suggest caching strategies. Reasoning models evaluate parallelization opportunities and dependency graphs.
Create deployment scripts, rollback procedures, and blue-green deployment configs. Function calling enables integration with cloud providers and registries.
Generate Terraform, Pulumi, and CloudFormation templates. Models understand resource dependencies, state management, and drift detection.