🏷️Topic

Gpt-5.4-mini

3 articles
First tracked: May 11, 2026
Last updated: Jul 30, 2026

Latest Coverage

Microsoft Says New Cybersecurity AI Model Helps MDASH Score 95.95% at Half the Cost

↗

Microsoft’s MDASH adds a cybersecurity-specific model (MAI-Cyber-1-Flash with GPT-5.4) that scores 95.95% on CyberGym and claims a 50% lower configuration cost than the existing MDASH stack.

Jul 30, 20261%

CVE-Bench: testing LLM agents on real-world vulnerability patches

↗

Five frontier LLMs were tested on 20 real CVEs across three prompt types; no model reliably fixes vulnerabilities, with a best 50% solve rate and significant cross-family differences; token cost varies up to ~4x by model, and locate prompts are the hardest test of genuine security reasoning.

May 29, 20261%

Interfaze: A new model architecture built for high accuracy at scale

↗

Interfaze markets a CNN/DNN+transformer hybrid architecture claiming high deterministic-task accuracy at scale, backed by benchmark leadership, large context, and competitive pricing.

May 11, 20261%

Related Entities

🏷️TopicMai-cyber-1-flash
4
📈StockMDASH
4
🏷️TopicGpt-5.3-codex
3
🏷️TopicGpt-5.4
13
🏷️TopicRestricted Access
1