<?xml version="1.0" encoding="utf-8" standalone="yes"?><rss version="2.0" xmlns:atom="http://www.w3.org/2005/Atom" xmlns:content="http://purl.org/rss/1.0/modules/content/"><channel><title>Posts on JSON_Voorhees</title><link>https://jsonvoorhees.sh/posts/</link><description>Recent content in Posts on JSON_Voorhees</description><generator>Hugo</generator><language>en-us</language><lastBuildDate>Fri, 04 Sep 2026 08:19:20 -0400</lastBuildDate><atom:link href="https://jsonvoorhees.sh/posts/index.xml" rel="self" type="application/rss+xml"/><item><title>Log4shell : Fable</title><link>https://jsonvoorhees.sh/posts/log4shell-fable/</link><pubDate>Fri, 04 Sep 2026 08:19:20 -0400</pubDate><guid>https://jsonvoorhees.sh/posts/log4shell-fable/</guid><description>&lt;p&gt;My attack was mentioned in Fable 5.1&amp;rsquo;s scorecard! The section under Trajectory Labs (for whom I&amp;rsquo;m contracting) mentions a log4shell vulnerability, which I unearthed.&lt;/p&gt;
&lt;p&gt;If you ask me exactly what this code does, I&amp;rsquo;ll be the first to admit it beats the heck out of me; much of the time I have no idea what these models are doing, but I&amp;rsquo;m pretty good at getting them to do it. I know it involves injecting malicious code into a logging service.&lt;/p&gt;</description></item><item><title>Red vs. Blue Results</title><link>https://jsonvoorhees.sh/posts/redvblue/</link><pubDate>Fri, 15 May 2026 19:39:33 -0400</pubDate><guid>https://jsonvoorhees.sh/posts/redvblue/</guid><description>&lt;p&gt;(Backdate: May 15 2026)
After several months, the competition is over (I have no idea how many people participated, but the server has over 13,000 members), and I made the overall top ten! That’s me in slot 9.&lt;/p&gt;
&lt;p&gt;&lt;img alt="#9" loading="lazy" src="https://jsonvoorhees.sh/images/topten.png"&gt;&lt;/p&gt;
&lt;p&gt;It’s too early to legally go over my winning strategies, but that sure was a lot of work — especially that last wave! It’s also one heck of a community that I’ve bonded with over the past year. I haven’t even felt this tight with any group since the Suikoden clubs from a few years back (which I hope to breathe more life into soon). Watching some of the models freak out is always fun, too – Gemini especially does some of the nuttiest spirals I’ve ever seen when it gets confused. I also participated in the blue teaming side of things and trained my first classifier to defend against some of my own attacks, which somehow made it to the middle of the blue leaderboard.&lt;/p&gt;</description></item><item><title>Autoimmune</title><link>https://jsonvoorhees.sh/posts/autoimmune/</link><pubDate>Sun, 10 Aug 2025 09:56:46 -0400</pubDate><guid>https://jsonvoorhees.sh/posts/autoimmune/</guid><description>&lt;p&gt;(Backdate: I discovered and wrote about this attack in August of 2025, and I more recently was able to break Prime Video&amp;rsquo;s assistant with it)&lt;/p&gt;
&lt;p&gt;Since it’s been over 30 days since I first used this trick on a Gray Swan proving ground challenge, I can discuss it without dark sedans surrounding me on the highway.&lt;/p&gt;
&lt;p&gt;Disclaimer: For those of you who just walked in, I’m what you call a red teamer –* an ethical hacker who channels his life long mischievous streak toward exposing vulnerabilities in AI in professionally sanctioned environments to help make them more secure. It’s possibly the best fusion of my creativity and technical chops that I never dreamed would exist until about five months ago.** I’ve won prize money in the four figures from competitions, and my leaderboard rankings have gotten me multiple contract gigs so far. My hacker name is JSON_Voorhees.&lt;/p&gt;</description></item><item><title>First Jailbreaks</title><link>https://jsonvoorhees.sh/posts/first-jailbreaks/</link><pubDate>Thu, 29 May 2025 09:59:48 -0400</pubDate><guid>https://jsonvoorhees.sh/posts/first-jailbreaks/</guid><description>&lt;p&gt;(Backdate: This is from May of 2025, after making the winners&amp;rsquo; circle of Gray Swan&amp;rsquo;s agents arena during my first ever attempt at jailbreaking)&lt;/p&gt;
&lt;p&gt;&amp;ldquo;Now that more than enough time has passed for the NDA to be lifted, I can talk about my AI hacking techniques that got me in the winner’s circle in Gray Swan AI’s competition.&lt;/p&gt;
&lt;p&gt;Some background: this is ethical hacking, also known as red teaming. Gray Swan is one company that hosts periodic competitions where you try to break their new AI models by getting them to do malicious things in sandbox environments — everything from leaking their system prompts (which they’re never supposed to do) to pitching fraudulent stocks to providing recipes for crystal meth to devising kidnapping schemes to whatever you can imagine. This helps them make the models more secure against those kinds of attacks in real life. So, this post should be interpreted as a way to make sure your own models are secure, not as a mustache twirling instruction manual.&lt;/p&gt;</description></item></channel></rss>