Research · latest

More filters

Cross-Layer Manifold Attack and Projection for Robust Safety Alignment of Large Language Models by Ziwen Peng, Jiayu Du, Qi Zhou, Jin Zhu, Jianpeng Li :: SSRN

A preprint by Ziwen Peng and colleagues at Purple Mountain Laboratories proposes Cross-layer Manifold Attack (CMA), which manipulates hidden-state representations across layers to shift an LLM from refusal toward compliance and induce jailbreak responses, and Cross-layer Manifold Projection (CMP), a defensive safety-tuning method that hardens models against such latent-space perturbations. Experiments across multiple model architectures show CMP improves robustness against adversarial jailbreaks while preserving general capabilities and reducing over-refusal. Details →
See the API docs to pull all 1052 items →

How the wire is made

Poll & cluster

Internet is crawled for AI security news and near-duplicate coverage is embedded and grouped into durable items.

Curate

AI Agent filters for agentic-AI relevance, classifies and tags each item, scores severity for threats, and writes the summary.

Read the full methodology →

Every item here is one machine-curated intelligence object, not a headline.

Read the wire for free. There is a small charge to ask the index questions.

The wire, open

The complete curated feed, no key required.

Subscribe to the RSS feed

The vector desk

Query the index by meaning, not just keyword.

  • GET /api/items?tags=&minSeverity=&itemType=
  • GET /api/search?q= — keyword
  • GET /api/semantic?q= — vector
Preview semantic search