Skip to main content
Models & Research

Mitsuba Squeezes a 27B Vision Model Into 7.3 GB on One GPU

Subtopic Small Models · Vision Language · Quantization Takeaways − Mitsuba is a ternary 1.58-bit quantization of Qwen3.8-27B, shrunk to 7.3 GB for a single 16 GB GPU. Purpose-built for ComfyUI: image to prompt generation for Stable Diffusion, Krea, and video pipelines.

By Precis Daily Newsroom2 min read360 words
Illustration for: Mitsuba Squeezes a 27B Vision Model Into 7.3 GB on One GPU
Illustration
Key points
  • Subtopic Small Models · Vision Language · Quantization Takeaways − Mitsuba is a ternary 1.58-bit quantization of Qwen3.8-27B, shrunk to 7.3 GB for a single 16 GB GPU.
  • Vision score 87.8 versus 89.8 for the BF16 original; coding collapsed from 66 to 4.
  • Mitsuba squeezes a 27B vision model into 7.3 GB Mitsuba is a ternary-quantized vision-language model designed to convert reference images into structured prompts for ComfyUI, Stable Diffusion, Krea, and similar generation tools.

Subtopic Small Models · Vision Language · Quantization Takeaways − Mitsuba is a ternary 1.58-bit quantization of Qwen3.8-27B, shrunk to 7.3 GB for a single 16 GB GPU. Purpose-built for ComfyUI: image to prompt generation for Stable Diffusion, Krea, and video pipelines. Vision score 87.8 versus 89.8 for the BF16 original; coding collapsed from 66 to 4. Optional 70 MB HiMitsuba LoRA cuts both outright refusals and evasive answers on sensitive requests. Requires the PrismML llama.cpp fork ; upstream does not yet support PQ2_0 or PTQ1_0. Must be run with thinking mode OFF or responses frequently come back empty. Mitsuba squeezes a 27B vision model into 7.3 GB Mitsuba is a ternary-quantized vision-language model designed to convert reference images into structured prompts for ComfyUI, Stable Diffusion, Krea, and similar generation tools. Its 27-billion-parameter base compresses to 7.3 GB or 6.0 GB, depending on the format, and the author reports that either build can run on a single 16 GB GPU. The release targets developers who need local image analysis and tightly constrained prompt writing. General chat, code generation, and long-document analysis fall outside its strengths. Mitsuba accepts an image and describes it in a form that another model can use for image or video generation. Its supported workflows include: Stable Diffusion prompts that follow word limits, include required terms, exclude forbidden terms, and end with a Negative line. Video prompts divided into intervals such as 0-3s , 3-6s , and 6-9s , with a camera movement assigned to each segment. Image descriptions covering objects, quantities, people, scenes, charts, and visible text. Prompt generation remains imperfect under strict constraints. The model card reports that Mitsuba satisfied every condition in six of ten test cases, so applications should validate required words, exclusions, length limits, and output structure before passing a response downstream. Mitsuba quantizes the official Qwen3.8-27B weights to roughly 1.58 bits per weight. Each weight is represented by one of three values: -1 , 0 , or +1 . The files use Prism ML’s GGUF quantization formats. You've reached the end of the free preview. Upgrade to AlphaSignal Pro to read the full article - and everything else behind the paywall.

Sources

Summarized from the linked originals.

Related stories

Illustration for: Runway Research | Introducing Praxis-1 - Runway
Models & Research

Runway app for iPhone Runway app for Android An open-weight world action model that turns Runway's video pretraining into control for real robots. “Pick up the tennis ball and put it in the box.” Today we're announcing Praxis-1 , our first open-weight world action model.

RunwayML Blog3 min
Illustration for: Kling 4.0 Debuts 30s AI Video Model - Briefs Finance
Models & Research

Warning : Undefined variable $stocks in /var/www/briefs.co/htdocs/wp-content/plugins/oxygen/component-framework/components/classes/code-block.class.php(133) : eval()'d code on line 448 Warning : foreach() argument must be of type array|object, null given in /var/www/briefs.co/htdocs/wp-content/plugins/oxygen/component-framework/components/classes/code-block.class.php(133) : eval()'d code on line 448 Warning : Undefined variable $funds in /var/www/briefs.co/htdocs/wp-content/plugins/oxygen/component-framework/components/classes/code-block.class.php(133) : eval()'d code on line 472 Warning : foreach() argument must be of type array|object, null given in /var/www/briefs.co/htdocs/wp-content/plugins/oxygen/component-framework/components/classes/code-block.class.php(133) : eval()'d...

Kling AI14 min