This paper, published on arXiv, introduces a new benchmark for evaluating open-set radio frequency fingerprinting systems. These systems identify wireless devices by their unique signal…
arXiv: Which Model Is Actually Serving You? IRIS: Budgeted Black-Box Auditing of Model Substitution and Routing Dilution in LLM Gateways
AI_SAFETY. Sourced from arxiv_cscr, summarised by Matproof.
AI Analysis
What changed and what to do.
This paper, published on arXiv, introduces a new auditing framework called IRIS designed to detect two specific risks in Large Language Model (LLM) gateways: model substitution and routing dilution. Model substitution occurs when a provider secretly swaps a requested high-cost, high-performance model for a cheaper, less capable one. Routing dilution happens when a gateway mixes outputs from multiple models without disclosure, degrading performance or introducing bias. The authors propose a budgeted, black-box method to test whether the model you are paying for is actually the one serving your requests.
The primary organizations affected are any entity deploying or procuring LLM services through third-party gateways, including cloud providers, AI startups, and enterprise IT departments. Financial services, healthcare, and legal sectors that rely on verifiable model outputs for compliance or liability reasons are especially vulnerable. Regulators and auditors monitoring AI safety and transparency under frameworks like the EU AI Act will also need to consider these risks.
Compliance teams should immediately review their contracts and service-level agreements with LLM providers to ensure explicit guarantees against model substitution and routing dilution. They should also begin evaluating whether to implement independent auditing tools like IRIS to verify model identity and output consistency. Finally, teams should document any discrepancies found and report them to relevant regulatory bodies, as undisclosed model changes may violate transparency obligations under emerging AI governance rules.
This summary is AI-generated for orientation purposes. For regulatory action, always consult the original source linked above.
More AI_SAFETY updates
Latest in AI_SAFETY.
This publication, titled "Unconditional Unclonable Encryption," introduces a theoretical cryptographic breakthrough that could render current encryption standards obsolete. The paper demonstrates a…
This publication introduces a new consensus mechanism extension called Themis, designed to mitigate Maximal Extractable Value (MEV) risks in application-specific blockchains. MEV refers to the profit…
This paper, published on arXiv, presents a new security hypothesis and formal model for cryptographically verifiable authorization of autonomous AI agents. It proposes a framework where AI agents can…
Map this to your controls
Connect regulatory changes to your compliance work.
Matproof maps every regulator update directly to your controls and surfaces the ones that affect your organisation — across 21 frameworks.