<?xml version="1.0" encoding="UTF-8"?><rss xmlns:dc="http://purl.org/dc/elements/1.1/" xmlns:content="http://purl.org/rss/1.0/modules/content/" xmlns:atom="http://www.w3.org/2005/Atom" version="2.0"><channel><title><![CDATA[Lizard inference engineering: Do not ship an unproven fast path]]></title><description><![CDATA[<p dir="auto"><strong>Inference Engineering · Day 7 · Evening</strong></p>
<p dir="auto"><img src="https://lizard-llm.qendryx.com/screenshots/Benchmark/Screenshot%202026-07-22%20213232.png" alt="Raw provider metrics — lizard-llm.qendryx.com" class=" img-fluid img-markdown" /></p>
<p dir="auto">We implemented an AVX-512 VNNI decode kernel and chose not to enable it by default.</p>
<p dir="auto">The instruction is promising on supported Intel and AMD CPUs. But on the hardware available for validation, the measured result remained inside ordinary run-to-run noise.</p>
<p dir="auto">Caterpillar ships the proven path as default and keeps VNNI behind an explicit experiment flag.</p>
<p dir="auto">Engineering credibility sometimes means declining to claim a speedup.</p>
<p dir="auto"><strong>Engineering fact:</strong> Caterpillar's AVX-512 VNNI kernel remains opt-in because it measured within normal run-to-run noise on available test hardware.</p>
<p dir="auto"><a href="https://lizard-llm.qendryx.com/benchmarks.html" rel="nofollow ugc">Read the relevant Lizard page</a></p>
<p dir="auto">#AVX512 #CPUOptimization #Benchmarking #CaterpillarEngine #LizardLLM</p>
<p dir="auto">&lt;!-- lizard-marketing-slot:day-07-pm --&gt;</p>
]]></description><link>https://community.lizard-llm.qendryx.com/topic/53/lizard-inference-engineering-do-not-ship-an-unproven-fast-path</link><generator>RSS for Node</generator><lastBuildDate>Mon, 24 Aug 2026 01:37:05 GMT</lastBuildDate><atom:link href="https://community.lizard-llm.qendryx.com/topic/53.rss" rel="self" type="application/rss+xml"/><pubDate>Thu, 30 Jul 2026 11:00:11 GMT</pubDate><ttl>60</ttl></channel></rss>