<?xml version="1.0" encoding="utf-8" standalone="yes"?><rss version="2.0" xmlns:atom="http://www.w3.org/2005/Atom"><channel><title>benchmark on Gaia Lab · Blog</title><link>https://blog.defectiv.es/en/tags/benchmark/</link><description>Recent content in benchmark on Gaia Lab · Blog</description><generator>Hugo</generator><language>en-GB</language><lastBuildDate>Wed, 23 Sep 2026 09:00:00 +0200</lastBuildDate><atom:link href="https://blog.defectiv.es/en/tags/benchmark/index.xml" rel="self" type="application/rss+xml"/><item><title>Measuring quantized LLMs: what is lost when an open model is compressed</title><link>https://blog.defectiv.es/en/posts/medir-llm-cuantizados-gguf-bfcl-bigcodebench-ruler/</link><pubDate>Tue, 22 Sep 2026 12:45:00 +0200</pubDate><guid>https://blog.defectiv.es/en/posts/medir-llm-cuantizados-gguf-bfcl-bigcodebench-ruler/</guid><description>&lt;p&gt;Third instalment of &lt;strong&gt;Cluster X-ray&lt;/strong&gt;. It is the best example of what the &lt;a href="https://blog.defectiv.es/en/posts/que-corre-de-verdad-en-nuestro-cluster/"&gt;overview post&lt;/a&gt;&#10; said: the cluster today is above all a &lt;strong&gt;measurement laboratory&lt;/strong&gt;. Not a single parameter is trained here.&lt;/p&gt;&#10;&lt;h2 id="the-question"&gt;The question &lt;a class="hanchor" href="#the-question" aria-label="Enlace a esta sección"&gt;#&lt;/a&gt;&lt;/h2&gt;&#10;&lt;p&gt;Open models are distributed quantized: the same weights compressed to 8, 6, 5, 4, 3 or 2 bits so they fit on a small card or a laptop. The community publishes hundreds of variants, but rarely with a comparable measure of what is lost. The line asks &lt;strong&gt;how well each quantization level of each family really performs on tasks that matter for using the model as a tool&lt;/strong&gt;, and at what speed and energy cost it does so on each card.&lt;/p&gt;</description></item></channel></rss>