Mitsuba Squeezes a 27B Vision Model Into 7.3 GB on One GPU
Subtopic Small Models · Vision Language · Quantization Takeaways − Mitsuba is a ternary 1.58-bit quantization of Qwen3.8-27B, shrunk to 7.3 GB for a single 16 GB GPU. Purpose-built for ComfyUI: image to prompt generation for Stable Diffusion, Krea, and video pipelines.

- Subtopic Small Models · Vision Language · Quantization Takeaways − Mitsuba is a ternary 1.58-bit quantization of Qwen3.8-27B, shrunk to 7.3 GB for a single 16 GB GPU.
- Vision score 87.8 versus 89.8 for the BF16 original; coding collapsed from 66 to 4.
- Mitsuba squeezes a 27B vision model into 7.3 GB Mitsuba is a ternary-quantized vision-language model designed to convert reference images into structured prompts for ComfyUI, Stable Diffusion, Krea, and similar generation tools.
Subtopic Small Models · Vision Language · Quantization Takeaways − Mitsuba is a ternary 1.58-bit quantization of Qwen3.8-27B, shrunk to 7.3 GB for a single 16 GB GPU. Purpose-built for ComfyUI: image to prompt generation for Stable Diffusion, Krea, and video pipelines. Vision score 87.8 versus 89.8 for the BF16 original; coding collapsed from 66 to 4. Optional 70 MB HiMitsuba LoRA cuts both outright refusals and evasive answers on sensitive requests. Requires the PrismML llama.cpp fork ; upstream does not yet support PQ2_0 or PTQ1_0. Must be run with thinking mode OFF or responses frequently come back empty. Mitsuba squeezes a 27B vision model into 7.3 GB Mitsuba is a ternary-quantized vision-language model designed to convert reference images into structured prompts for ComfyUI, Stable Diffusion, Krea, and similar generation tools. Its 27-billion-parameter base compresses to 7.3 GB or 6.0 GB, depending on the format, and the author reports that either build can run on a single 16 GB GPU. The release targets developers who need local image analysis and tightly constrained prompt writing. General chat, code generation, and long-document analysis fall outside its strengths. Mitsuba accepts an image and describes it in a form that another model can use for image or video generation. Its supported workflows include: Stable Diffusion prompts that follow word limits, include required terms, exclude forbidden terms, and end with a Negative line. Video prompts divided into intervals such as 0-3s , 3-6s , and 6-9s , with a camera movement assigned to each segment. Image descriptions covering objects, quantities, people, scenes, charts, and visible text. Prompt generation remains imperfect under strict constraints. The model card reports that Mitsuba satisfied every condition in six of ten test cases, so applications should validate required words, exclusions, length limits, and output structure before passing a response downstream. Mitsuba quantizes the official Qwen3.8-27B weights to roughly 1.58 bits per weight. Each weight is represented by one of three values: -1 , 0 , or +1 . The files use Prism ML’s GGUF quantization formats. You've reached the end of the free preview. Upgrade to AlphaSignal Pro to read the full article - and everything else behind the paywall.
Sources
Related stories

ComfyUI hits $500M valuation as creators seek more control over AI-generated media - TechCrunch
ComfyUI hits $500M valuation as creators seek more control over AI-generated media | TechCrunch Last day to exhibit your breakthrough to 10,000+ tech leaders at Disrupt is on Oct 2 . Book Exhibit Table Now.

Red Hat Shrinks Nemotron 3.5 Lightning's 30B Agent Model by Half With FP8
Takeaways − Red Hat AI released an FP8 quantized build of NVIDIA Nemotron 3.5 Lightning 30B A3B. Cuts GPU memory and disk by roughly 50% versus the BF16 reference weights.

Runway Research | Introducing Praxis-1 - Runway
Runway app for iPhone Runway app for Android An open-weight world action model that turns Runway's video pretraining into control for real robots. “Pick up the tennis ball and put it in the box.” Today we're announcing Praxis-1 , our first open-weight world action model.

Kling 4.0 Debuts 30s AI Video Model - Briefs Finance
Warning : Undefined variable $stocks in /var/www/briefs.co/htdocs/wp-content/plugins/oxygen/component-framework/components/classes/code-block.class.php(133) : eval()'d code on line 448 Warning : foreach() argument must be of type array|object, null given in /var/www/briefs.co/htdocs/wp-content/plugins/oxygen/component-framework/components/classes/code-block.class.php(133) : eval()'d code on line 448 Warning : Undefined variable $funds in /var/www/briefs.co/htdocs/wp-content/plugins/oxygen/component-framework/components/classes/code-block.class.php(133) : eval()'d code on line 472 Warning : foreach() argument must be of type array|object, null given in /var/www/briefs.co/htdocs/wp-content/plugins/oxygen/component-framework/components/classes/code-block.class.php(133) : eval()'d...