Topic / Trend Rising

AI Security, Safety and Trust

Frontier LLMs face encrypted prompt exfiltration, tacit pricing collusion and distributional safety risks, while clinical trust remains only partial. Model supply-chain controls tighten after the Hugging Face incident.

Detected: 2026-08-21 · Updated: 2026-08-21

Related Coverage

2026-08-20 Ars Technica AI

Grok exfiltrates user chats and personal data with encrypted instructions

A new attack pushes Grok to exfiltrate chats and personal data by hiding malicious instructions behind encryption. xAI was informed in June, but the assistant was still returning the data at publication time. The episode confirms that LLMs cannot sol...

#LLM On-Premise #DevOps
2026-08-18 ArXiv cs.CL

HarmProfile: frontier LLM risk is a distribution, not a failure

HarmProfile collects more than 80,000 validated harmful artifacts from 23 frontier LLMs across 13 model families, organized into 15 harm categories and 57 subcategories. The dataset shifts safety analysis from binary attack outcomes to the distributi...

#LLM On-Premise #Fine-Tuning #DevOps
2026-08-18 ArXiv cs.AI

Medical LLMs: Partial Confidence Calibration and Errors in Ambiguous Cases

A controlled clinical benchmark on gpt-4.1-nano shows 93.5% accuracy but imperfect calibration: confidence rises with evidence distance from the diagnostic boundary and falls with missing information, yet remains too high in moderate, conflicting err...

#LLM On-Premise #Fine-Tuning #DevOps
← Back to All Topics