Security Sonar

AI security research and advisory.

Read the Research contact@securitysonar.com

Latest Research

Benchmarking the Security of Open-Weight Models, Part 4: What a Detection Standard Catches, and Why
Benchmarking the Security of Open-Weight Models, Part 4: What a Detection Standard Catches, and Why

Part 3 showed that a harness’s own guardrail can’t see inside file content it reads. This installment layers a product...

Benchmarking the Security of Open-Weight Models, Part 3: The Harness Around the Model
Benchmarking the Security of Open-Weight Models, Part 3: The Harness Around the Model

Parts 1 and 2 benchmarked whether Qwen 3.8 is safe to deploy. This piece asks a different question: once that same model is wired ...

Benchmarking the Security of Open-Weight Models, Part 2: A Local Stack on DGX Spark (with Qwen 3.8 as a Case Study)
Benchmarking the Security of Open-Weight Models, Part 2: A Local Stack on DGX Spark (with Qwen 3.8 as a Case Study)

Part 1 made the case that security benchmarking is a discipline distinct from capability benchmarking. This installment puts it in...