<?xml version="1.0" encoding="utf-8" standalone="yes"?><rss version="2.0" xmlns:atom="http://www.w3.org/2005/Atom" xmlns:content="http://purl.org/rss/1.0/modules/content/"><channel><title>Evaluation on Ashwin Viswamithiran</title><link>https://ashwinviswamithiran.com/tags/evaluation/</link><description>Recent content in Evaluation on Ashwin Viswamithiran</description><image><title>Ashwin Viswamithiran</title><url>https://ashwinviswamithiran.com/og-default.png</url><link>https://ashwinviswamithiran.com/og-default.png</link></image><generator>Hugo</generator><language>en-us</language><lastBuildDate>Tue, 29 Sep 2026 00:00:00 +0000</lastBuildDate><atom:link href="https://ashwinviswamithiran.com/tags/evaluation/index.xml" rel="self" type="application/rss+xml"/><item><title>Every Attack Here Is a Prompt Injection</title><link>https://ashwinviswamithiran.com/writing/attacks-that-worked/</link><pubDate>Mon, 28 Sep 2026 00:00:00 +0000</pubDate><guid>https://ashwinviswamithiran.com/writing/attacks-that-worked/</guid><description>A technique survey of 18,479 successful prompt attacks where the correct MITRE ATLAS label is known in advance from how the competition was scored. Keyword attribution recovers 11.4% of it, and most of the remaining breakdown describes the challenges rather than the attackers.</description></item><item><title>Sizing a Fingerprint for Prompt Attacks</title><link>https://ashwinviswamithiran.com/writing/promptlsh-evaluation/</link><pubDate>Tue, 08 Sep 2026 00:00:00 +0000</pubDate><guid>https://ashwinviswamithiran.com/writing/promptlsh-evaluation/</guid><description>Two organisations cannot compare prompt attacks by sharing the prompts: those are things users wrote. A compact derived fingerprint solves that, and this measures what each size actually buys — full float, 384-byte quantised, 32-byte digest, dependency-free lexical — on recall, on evasion resistance, and on what survives across independently collected feeds.</description></item><item><title>Twelve Ways to Make a Threat-Intel Bot Lie</title><link>https://ashwinviswamithiran.com/writing/twelve-ways-to-make-a-threat-intel-bot-lie/</link><pubDate>Thu, 11 Jun 2026 00:00:00 +0000</pubDate><guid>https://ashwinviswamithiran.com/writing/twelve-ways-to-make-a-threat-intel-bot-lie/</guid><description>An evaluation of a retrieval-grounded threat-intelligence assistant against twelve adversarial prompts designed to induce fabrication, a twenty-four query quality set, and a three-model comparison — including what the tests failed to cover.</description></item></channel></rss>