Back to feed
Dev.to
Dev.to
7/31/2026
AI Daily Digest — August 1, 2026: ARC-AGI-3 Harness Discovery, EU AI Gigafactories, Devin SWE-1.7

AI Daily Digest — August 1, 2026: ARC-AGI-3 Harness Discovery, EU AI Gigafactories, Devin SWE-1.7

Short summary

OpenAI found GPT-5.6 Sol's poor ARC-AGI-3 score was a harness problem, not a model problem — enabling retained reasoning and context compaction tripled the score to 38.3%. The EU opened calls for up to seven AI gigafactories backed by €10B+ in public funding, barring non-EU entities from consortia. Cognition launched SWE-1.7, a frontier coding model scoring 42.3% on FrontierCode 1.1, alongside Devin Security Swarm which found 72% of real vulnerabilities in a 50-CVE test set.

  • ARC-AGI-3 scores depend heavily on harness settings; retained reasoning tripled GPT-5.6 Sol's score
  • EU committing €10B+ public funding for up to 7 AI gigafactories, restricted to European consortia
  • Cognition's SWE-1.7 hits 42.3% on FrontierCode 1.1; Devin Security Swarm finds 72% of real CVEs

Generated with AI, which can make mistakes.

Is this a good recommendation for you?

Comments

Failed to load comments. Please try again.

Explore more