First reported arxiv.org
Research · latest
First reported snyk.io
Snyk VulnBench JS 1.0: LLM Bug Repeatability
Snyk VulnBench JS 1.0 is a benchmark study that ran 300 repeated vulnerability-finding scans to measure how repeatable an agentic LLM security review is on identical code, prompt, and harness. It found LLM findings unevenly repeatable: reference-matched findings were stable while extra-model reports varied widely, with nearly 50% of LLM-only reports appearing in just one of five identical scans, and the best LLM configuration reaching only 75.4% F1 against deterministic SAST. Details →How the wire is made
Poll & cluster
Internet is crawled for AI security news and near-duplicate coverage is embedded and grouped into durable items.
Curate
AI Agent filters for agentic-AI relevance, classifies and tags each item, scores severity for threats, and writes the summary.
Every item here is one machine-curated intelligence object, not a headline.
Read the wire for free. There is a small charge to ask the index questions.
The wire, open
The complete curated feed, no key required.
- GET /feed.xml — RSS 2.0, every item
- GET /api/items — read-only
The vector desk
Query the index by meaning, not just keyword.
- GET /api/items?tags=&minSeverity=&itemType=
- GET /api/search?q= — keyword
- GET /api/semantic?q= — vector