CESBench Introduces New Benchmark for Testing LLMs on IoT Cryptographic Security
Researchers create 380-item benchmark evaluating large language models on cryptographic engineering security for IoT devices.
Daily AI briefing
A fast, sourced read on the companies, research, infrastructure, and policy moves shaping artificial intelligence.
Lead story CESBench Introduces New Benchmark for Testing LLMs on IoT Cryptographic Security Sep 22Today in AI
Researchers create 380-item benchmark evaluating large language models on cryptographic engineering security for IoT devices.
Investigation documents cases of migrants who died after passing undetected through areas monitored by AI-enabled surveillance towers along the US-Mexico border.
OpenAI announces independent advisory group at IAS as new AI model solves 100+ mathematics problems, drawing mixed response from mathematicians.
A new large language model aims to fill gaps in fragmented ancient Greek papyrus records to reveal details about ancient life.
New research tackles conversational memory systems, encrypted guardrails against jailbreaks, and AI tutoring for programming students.
Longer horizon
How Cursor's massive November 2025 funding round and $1B+ revenue run rate signaled the mainstreaming of AI-powered software development.
In November 2025, Anthropic released Claude Opus 4.5, its most powerful model yet, advancing reasoning and coding capabilities with enhanced safety features.
How Anthropic's massive November 2025 infrastructure commitment signaled a new phase in enterprise AI competition and domestic manufacturing.