
Post-training image models for fandom
The best open-source image models make great pictures. Character.ai needed more: the same Character, recognizable through every style, scene, and page a story…
Open weights, libraries, datasets and community projects.

The best open-source image models make great pictures. Character.ai needed more: the same Character, recognizable through every style, scene, and page a story…

Migrations are the bane of any mature company. Systems are deeply integrated, you have key stakeholders across domains, and a small improvement can take months or years to integrate in traditional cases.

Hey folks,We had some reports of ending up in spam 😭 - so if you add us as a contact in Gmail, that shouldn’t happen.You’ll have seen me…

Our 255th episode with a summary and discussion of last week’s big AI news!Recorded on 08/26/2026Hosted by Andrey Kurenkov and Jeremie HarrisFeel free…

Ollama's Pro, Max, and Team plans now use transparent per-token pricing. Based on your feedback, every plan includes a monthly pool of usage credits. If you're on an existing Pro, Max, or Team plan, your plan continues to work as-is.

Weaviate v1.39 is now available open-source and on Weaviate Cloud. Two search features reach general availability in this release: the Boost API for…

Welcome to Import AI, a newsletter about AI research. Import AI runs on arXiv, cappuccinos, and feedback from readers.

NVIDIA Nemotron 3.5 Lightning is now available on Ollama, and it runs completely on your own device. It's a 30 billion parameter (3B active) open model from NVIDIA built for agents that stay running: gathering context, calling tools, and working through multi-step tasks.

Muse Glimmer from Meta Superintelligence Labs is now available Muse Glimmer , Meta's newest open model and the first released by Meta Superintelligence Labs, is now available on Ollama . It's a 30B multimodal model purpose-built for agent workloads that run locally with a 128K+ context length, released under the Apache 2.0 license.

At a glance Orchard is an open-source framework for scalable and cost-effective agentic AI research, built around Orchard Env, a reusable environment service…

Note from Andrey: apologies for the newsletter not having resumed yet - the next release is going to come tomorrow, and it will resume weekly cadence…

Michael and I first met in college, where we started our first company, Kitematic, which made Docker dead-simple to run. In 2015, it was acquired by Docker.

Faster Gemma 4 on MLX with multi-token prediction Gemma 4 is now significantly faster in Ollama 0.31. On Apple Silicon, it generates tokens nearly 90% faster on average across a coding-agent benchmark.
Ollama's highest performance on Apple Silicon yet with MLX Ollama's MLX engine has been updated to deliver its highest performance on Apple Silicon yet. By leaning more heavily on Apple's unified memory and the Metal-backed MLX framework, models output higher quality responses, respond faster, and use less memory.

Improved performance and model support with GGUF Ollama 0.30 is now available with improved performance and GGUF model compatibility through llama.cpp . This augments Ollama's MLX engine on Apple silicon, bringing support to more models on a wider range of hardware.

NVIDIA Nemotron 3 Ultra is now available on Ollama’s cloud. It’s a 550 billion parameter (55B active) open model from NVIDIA built for long-running, agentic workflows with fast and affordable performance across hundreds of tool calls.

Many people asked me over the past months to share my workflow for how I come up with the LLM architecture sketches and drawings in my articles, talks, and…

We are excited to introduce Qwen3Guard, the first safety guardrail model in the Qwen family. Built upon the powerful Qwen3 foundation models and fine-tuned specifically for safety classificatoin, Qwen3Guard ensures responsible AI interactions by delivering precise safety detection for both prompts and responses, complete with risk levels and categorized classifications for accurate moderation.

We are excited to introduce Qwen-Image-Edit, the image editing version of Qwen-Image. Built upon our 20B Qwen-Image model, Qwen-Image-Edit successfully extends Qwen-Image's unique text rendering capabilities to image editing tasks, enabling precise text editing.

We are thrilled to release Qwen-Image , a 20B MMDiT image foundation model that achieves significant advances in complex text rendering and precise image editing. To try the latest model, feel free to visit Qwen Chat and choose “Image Generation”.

Today, we're announcing Qwen3-Coder, our most agentic code model to date.

We release Qwen3 Embedding series , a new proprietary model of the Qwen model family. These models are specifically designed for text embedding , retrieval , and reranking tasks, built on the Qwen3 foundation model.

Today, we are excited to announce the release of Qwen3 , the latest addition to the Qwen family of large language models. Our flagship model, Qwen3-235B-A22B , achieves competitive results in benchmark evaluations of coding, math, general capabilities, etc., when compared to other top-tier models such as DeepSeek-R1, o1, o3-mini, Grok-3, and Gemini-2.5-Pro.