Cur8 — Wednesday, September 02, 2026
Europe's Quasar 438B model tops charts, signaling strong AI progress outside US. BenchMIRT questions LLM benchmark validity, urging reevaluation of assessment methods. Mistral's default user input training raises privacy concerns, except for enterprise users. OpenAI expands healthcare data integration and shares AI governance insights from top law firm. Anthropic updates Claude models with enhanced creativity and safety features. Google introduces agentic video understanding and easy image tools in Workspace. NYC bans AI use for students up to 8th grade, setting precedent for education policy.
New model releases
Quasar 438B, Europe's top AI model, scores 43 on the Artificial Analysis Intelligence Index, outperforming competitors like Mistral Medium 3.5 (30) and NVIDIA Nemotron 3 Ultra (38). It processes 500 tokens in 15.3 seconds, excelling in long-context reasoning (75.0) and terminal work (69.3), available via CompactifAI API.
BenchMIRT method audits LLM benchmarks at individual prompt level, revealing mixed signals in single scores; trained on 100 LLMs across 16 benchmarks, it identified safety and general reasoning as dominant dimensions.
Open weights models
Hardware limits local coding LLM use; top models require APIs (GLM 5.3: 743B params, 66.9% Deep SWE; Kimi K 2.7: 1T params, agentic tasks). Local options exist for varying hardware (Qwen 3 Code Next: 80B params, 64GB+ Mac/AMD; Qwen 3.6 27B: 27B params, 24GB GPU). Codestral 2 offers local tab completion rivaling GitHub Copilot.
Mistral now uses user input for model training by default, except for enterprise tier users. Users can opt out via admin panel or app settings.
OpenAI news
OpenAI's Astra model meets Critical cybersecurity capability threshold with enhanced safeguards for release.
Apple claims forensic analysis of ex-employee's MacBook reveals trade secret misuse at OpenAI; seeks expedited discovery. Evidence includes use of confidential schematics in AI agent training.
AI-native companies leverage AI agents for enhanced onboarding, account management, and developer integrations.
Healthcare orgs integrate EHR & industry data with ChatGPT for secure clinician access to patient context & medical research.
Gilbert + Tobin scaled ChatGPT Enterprise and Codex firm-wide through CEO-led commitment, rigorous governance, and human accountability.
Anthropic news
Claude Fable 5.1 and Mythos 5.1 released: 25-45% cheaper, better benchmarks, new safeguards; Mythos restricted access. Fable 5.1 excels in coding (73.4% CursorBench), science (52.6% Terminal-Bench-Science), and knowledge work (1853 GDPval-AA). Mythos 5.1 designs high-affinity binders, speeds up deep learning models. Enterprise Frontier Safeguards coming fall.
Claude Fable 5.1 sets coding, knowledge work standard; scores 52.6% on Terminal-Bench-Science 0.1. Max reasoning effort generates detailed SVG pelican for $3.30 in 14 minutes.
Claude 5.1 now refuses to reproduce song lyrics, copyrighted characters, or logos; provides harm-reduction info but not drug protocols; and avoids apologizing to rude users. Fable's reliable knowledge cutoff is June 2026.
Claude Fable 5.1 launched on AWS, excelling in coding, research, and enterprise tasks; Enterprise Frontier Safeguards offer zero data retention for eligible customers through 2026. Available now in select regions, with pricing detailed on Amazon Bedrock.
Big cloud & vendor AI news
Gemini 3.7 Flash introduced at half the cost of 3.6 Flash; Pixel 11 series launched with Gemini Intelligence; Gemini app hit 1 billion monthly users; Gemma reached 1 billion downloads; WeatherNext 2 open-sourced for cyclone prediction.
Google Pics, built on Nano Banana model, rolls out to AI Pro/Ultra subscribers & Workspace business customers. Offers object segmentation, in-image text editing, collaboration, and multiple generations; integrates with Docs, Slides, Drive.
Agentic video understanding reduces token usage by up to 88%, costs by 66%, boosts accuracy by 7% in Gemini models; available now via API.
Efficient frontier in LLM inference balances latency, throughput, cost; techniques target specific points or push entire frontier out. Tradeoff management includes batch sizing, parallelism strategies, quantization; frontier-moving techniques include kernel optimization, speculative decoding, disaggregation.
YouTube trending AI content
NYC bans AI use for students up to 8th grade, effective immediately.