Originally published on MakerPulse.
4.61 versus 4.55.
That's the gap between the top two models in our first AgentPulse benchmark run: GPT-5.2 and Gemini 3.1 Pro, separated by six hundredths of a point on task quality, scored by three independent AI evaluators across 28 real-world prompts. One costs $0.74 to run the full suite. The other...
🛡️ VERIFIED CYBER INTELLIGENCE ID: #3282844