AI News · 2026-07-31

AI News · 2026-07-31
💡

Jason Says

Today's most archivable signal: GPT-5.6 triples ARC-AGI-3 scores with just two API settings unchanged — this isn't benchmark gaming, it's proof that *how you call the model* is your real competitive moat, not which model you pick.

🛠️
AI ToolsOpenAI Blog

Two API Settings Triple GPT-5.6 ARC-AGI-3 Benchmark Scores

OpenAI reveals that enabling reasoning retention and compaction in the GPT-5.6 API tripled scores on ARC-AGI-3. This is a calling-pattern upgrade, not a new model — the same model at the same cost performs dramatically better, giving developers an immediate and free performance lever.

🛠️
AI ToolsOpenAI Blog

GPT-5.6 Luna Drops 80% in Price, Threatening Gemini Flash

OpenAI cuts GPT-5.6 Luna pricing by 80% and Terra by 20%, making Luna a direct competitor to Gemini Flash 3.1. For developers running high-volume agentic workflows, this dramatically lowers inference costs. Sol remains unchanged and still leads on quality, creating a clearer tiered pricing structure.

🛠️
AI ToolsTechCrunch AI

Google's AI Fixed More Chrome Bugs in June Than the Past Two Years Combined

Google reports that AI-assisted tools helped discover and patch more Chrome bugs in a single month (June 2026) than the previous two years combined. Following similar results at Microsoft, this data-backed proof point confirms AI-driven security is no longer experimental — it's production-grade engineering.

🛠️
AI ToolsTechCrunch AI

Okta Acquires AI Security Startup Permiso for ~$200M

Identity giant Okta is acquiring AI security startup Permiso for approximately $200M. Permiso specializes in detecting threats from AI agents and non-human identities (NHI) across cloud environments. As enterprises deploy agents at scale, NHI security is becoming critical infrastructure — and Okta just bought the key.

📚
AI PapersHuggingFace Papers

Comprehensive Survey: Memory Architectures for Large Language Models

This survey systematically maps LLM memory into four architectural dimensions: transient attention, recurrent state, parameter-efficient adaptation, and scalable lookup storage. For developers debating RAG vs fine-tuning vs long context, this framework provides a principled decision guide — not academic taxonomy for its own sake.

💰AI Funding Roundup

Inforcer

Series C$50M

Investors: Insight Partners

帮助中小企业应对 AI 时代安全合规风险的伦敦初创,Insight Partners 领投,填补 SMB 市场 AI 安全工具空白。

Subscribe for daily AI updates + free playbook

📘 Subscribe Free