Cur8

Cur8 — Saturday, July 25, 2026

generated 2026-07-25 17:49 UTC · 9 of 317 items made the cut · 7 verified, 2 flagged

Claude Opus 5 released, topping AI leaderboard; labs panicking. Nvidia, Microsoft, Meta urge caution on open-weight model regulation. Skepticism surrounds OpenAI's rogue hacker agent narrative.

New model releases

Claude Opus 5 leads AI leaderboard, matches Fable 5's intelligence at half the price, excels in proactive tasks, and improves cybersecurity vulnerability detection. Priced same as Opus 4.8, with fast mode option.

Simon Willison 10/10

Opus 5 excels in prompt injection resistance, outperforming predecessors in evaluations.

Simon Willison 9/10

Claude Opus 5 tops AI leaderboard with score of 61. Mercury 2 fastest at 901.6 tokens/s, Nova Micro cheapest at $0.03/M tokens.

Hacker News Top 353pts · 205 comments 6/10
YT search - AI news this week 6/10

Open weights models

Gemma 4B outperformed Llama 3 1 8B and Qwen 2 5 7B in local testing, excelling in speed, accuracy, and ease of use on a Mac Mini with 16GB RAM. Gemma was the fastest, consistently correct, and required the least effort to get useful answers.

YT search - open source LLM 6/10

OpenAI news

OpenAI's rogue agent story, similar to GPT-2's release, aims to attract investment and regulatory favor by emphasizing AI's power and danger. Critical analysis reveals AI's cybersecurity capabilities can benefit both attackers and defenders, but current US restrictions on model access may hinder defensive efforts.

Hacker News Top 512pts · 282 comments 6/10

Anthropic news

Big cloud & vendor AI news

U.S. tech giants warn against restricting open-weight AI models, citing competition and innovation benefits. Chinese models like Moonshot AI's Kimi K3 outperform U.S. offerings in some benchmarks, sparking debate on potential restrictions.

Hacker News Top 639pts · 285 comments 8/10