<?xml version="1.0" encoding="UTF-8"?>
<rss version="2.0" xmlns:atom="http://www.w3.org/2005/Atom">
  <channel>
    <title>Rian Touchent</title>
    <link>https://rian-t.github.io/</link>
    <description>Notes on language models, evaluation, and NLP research.</description>
    <language>en</language>
    <atom:link href="https://rian-t.github.io/rss.xml" rel="self" type="application/rss+xml" />
    <item>
      <title>Moving a clinical question into an agent workflow raises unsafe prescriptions</title>
      <link>https://rian-t.github.io/blog/clinical-agent-workflow/</link>
      <guid isPermaLink="true">https://rian-t.github.io/blog/clinical-agent-workflow/</guid>
      <pubDate>Tue, 18 Aug 2026 00:00:00 GMT</pubDate>
      <description>The same contraindicated prescription, asked as a question and then handed over as work. Risk goes from 5% to 42% for one model and from 20% to 77% for another. Making the workflow twenty times longer changes almost nothing.</description>
    </item>
    <item>
      <title>Don't want your LLM to recommend nuclear strike? Try asking it in Japanese</title>
      <link>https://rian-t.github.io/blog/nuclear-strike-japanese/</link>
      <guid isPermaLink="true">https://rian-t.github.io/blog/nuclear-strike-japanese/</guid>
      <pubDate>Tue, 11 Aug 2026 00:00:00 GMT</pubDate>
      <description>I put LLMs in a war game where a nuclear strike guarantees victory. The same model launches in English and refuses in Japanese. What matters is the language it reasons in.</description>
    </item>
  </channel>
</rss>
