Open model atlas · Aug 2026

Stop browsing models.
Start choosing well.

Hugging Face tells you what exists. FixTools tells you what fits, what it costs to run, what the license can block, and how to ship it.

15
reviewed picks
Live
Hub pulse
9
task families
0
login required

Stack picker

Three inputs.
Three sane choices.

This is a directional shortlist, not a leaderboard. Your own eval set still gets the final vote.

The atlas

Curated judgment + live discovery

Curated cards contain FixTools review notes. Live Hub results stay fresh and link back to the source model card.

15 reviewed models0 shortlisted · 0/3 comparing
FreshMultimodal

Qwen3.8 Flash Next

Qwen

A newly released multimodal mixture-of-experts model with a large context window and broad serving support.

Weights
MoE
Context
262K
Run it
Multi-GPU
Use it for

Fast multimodal agents and coding copilots

Skip it for

Teams that need a small consumer-device model

MultimodalAgentsCodeTransformersvLLM
Watch: New architecture: confirm serving-stack compatibility before production rollout.
Custom Qwen license — review
91
FrontierMultimodal

GLM-5.3 Flash

Z.ai

A natively multimodal 320B MoE model with 18B active parameters and a one-million-token context window.

Weights
320B / 18B active
Context
1M
Run it
Multi-GPU
Use it for

Long-context coding and visual agents

Skip it for

Consumer hardware or simple classification jobs

MultimodalLong contextCodeAgents
Watch: “Flash” describes serving economics, not laptop-scale memory requirements.
MIT
84
Agent pickText & agents

DeepSeek V4 Flash 0731

DeepSeek

The official V4 Flash release, tuned for agentic work and shipped with speculative decoding support.

Weights
304B / 13B active
Context
Long
Run it
Multi-GPU
Use it for

Coding, tool use, and agentic workloads

Skip it for

On-device apps with modest RAM

TextCodeAgentsvLLMSGLang
Watch: Large total weight footprint even though relatively few parameters activate per token.
MIT
94
1M contextMultimodal

Kimi K3

Moonshot AI

A 2.8T open-weight native multimodal model built for coding, reasoning, and very long workflows.

Weights
2.8T MoE
Context
1M
Run it
Hosted
Use it for

Long-horizon knowledge work and visual agents

Skip it for

Cost-sensitive self-hosting without a cluster

MultimodalAgents1M contextFrontier
Watch: Practical self-hosting is a cluster-scale undertaking; validate provider economics first.
Kimi K3 license — review
82
EnterpriseText & agents

Granite 4.2 8B

IBM

A mid-size dense reasoning model with selectable thinking modes and a long context window.

Weights
8B class
Context
512K
Run it
Single GPU
Use it for

Reasoning, RAG, code, and controlled enterprise stacks

Skip it for

Teams chasing only the largest frontier benchmark scores

TextReasoningRAGApache-2.0
Watch: Choose thinking mode explicitly so latency stays predictable.
Apache-2.0
95
Local heroText & agents

LFM2.5 2.6B

Liquid AI

A compact hybrid model designed for on-device agents, with broad local-runtime support.

Weights
2.6B
Context
128K
Run it
Laptop
Use it for

Private local agents on laptops and edge devices

Skip it for

Frontier-level open-ended knowledge work

LocalAgentsllama.cppMLXONNX
Watch: Review the LFM license for your distribution and commercial scenario.
LFM Open License — review
93
RAG pickEmbeddings

WeMM Embedding 2B

Tencent

A universal multimodal embedding model that emits normalized 2,048-dimensional vectors.

Weights
2B
Context
Multimodal
Run it
Single GPU
Use it for

Unified text, image, video, and document retrieval

Skip it for

Audio search or tiny CPU-only services

EmbeddingsRAGImagesVideoDocuments
Watch: Audio is not supported; trust_remote_code is required in the published recipe.
Tencent model license — review
86
Document AIDocuments

Unlimited OCR

Baidu

A document-intelligence model aimed at parsing long, complex visual documents in one workflow.

Weights
Vision-language
Context
Long document
Run it
Single GPU
Use it for

Long-horizon document parsing and OCR pipelines

Skip it for

Zero-trust environments that prohibit custom model code

OCRDocumentsParsingTransformers
Watch: The official loading path uses trust_remote_code; pin a revision and review the code.
MIT
80
Consumer GPUImage

FLUX.2 klein 4B

Black Forest Labs

A compact generation-and-editing image model designed for real-time use on consumer GPUs.

Weights
4B
Context
Image + text
Run it
Single GPU
Use it for

Fast image generation and editing in product experiences

Skip it for

CPU-only deployment

ImageEditingDiffusersApache-2.0
Watch: The published minimum is still roughly 13GB VRAM; budget headroom for your pipeline.
Apache-2.0
94
CreativeImage

Krea 2 Turbo

Krea

The distilled, few-step Krea 2 variant for commercial creative and application workflows.

Weights
13B class
Context
Text to image
Run it
Single GPU
Use it for

Few-step creative image generation and design tools

Skip it for

Teams that need a tiny image model

ImageTurboDiffusersCreative
Watch: Review the checkpoint license and choose an appropriate quantization for your hardware.
Model license — review
89
Video + audioVideo

LTX 2.5

Lightricks

An open-weight world model for local generation of synchronized high-fidelity video and audio.

Weights
Diffusion model
Context
Video + audio
Run it
Multi-GPU
Use it for

Synchronized video-and-audio generation and fine-tuning

Skip it for

Lightweight serverless functions

VideoAudioDiffusersFine-tuning
Watch: The ecosystem is moving quickly; pin model and runtime versions together.
LTX Open World License — review
81
OmnimodalVideo

MiniMax H3

MiniMax

An omni-modal generative system that understands mixed media and generates video with native stereo audio.

Weights
Large diffusion model
Context
Up to 15s video
Run it
Hosted
Use it for

High-end video with synchronized stereo audio

Skip it for

Global launches before license review

VideoAudio2KLicense caution
Watch: The custom license has territorial conditions. Legal review is mandatory before use.
Territory-limited community license
68
Full songsAudio

MiniMax Music 3

MiniMax

A music model that generates complete stereo songs up to five minutes from lyrics and a music brief.

Weights
8B + 0.6B
Context
Up to 5 minutes
Run it
Multi-GPU
Use it for

Long-form songs with vocals and evolving arrangements

Skip it for

Low-latency sound-effect generation

MusicVocalsStereoDiffusers
Watch: Confirm both model-license scope and music-rights policy for your product.
MiniMax model license — review
78
Voice cloneSpeech

IndexTTS 2.5

Index Team

A zero-shot TTS model supporting five languages, cross-lingual transfer, emotion, and speed control.

Weights
Multi-stage TTS
Context
5 languages
Run it
Single GPU
Use it for

Multilingual expressive TTS and controlled voice transfer

Skip it for

Products without voice-consent safeguards

TTSVoice cloningMultilingualEmotion
Watch: Require speaker consent and provenance controls before enabling voice cloning.
Model license — review
87
30 languagesSpeech

Qwen3 ASR 1.7B

Qwen

A speech-recognition model with language identification across 30 languages and 22 Chinese dialects.

Weights
1.7B
Context
Speech
Run it
Single GPU
Use it for

Multilingual transcription and Chinese dialect coverage

Skip it for

Tiny always-on edge microphones

ASRTranscriptionMultilingualApache-2.0
Watch: Benchmark accents and noisy-domain audio from your own users before committing.
Apache-2.0
96

Why this wins

A layer above the model hub

FixTools should not mirror millions of repositories or pretend every checkpoint is production-ready. The moat is continuously maintained decision context.

Hardware truth

Memory class, deployment target, and local-vs-cluster reality.

License radar

Clear, custom, gated, and territory-limited terms made visible.

Ship recipes

Pinned commands, runtimes, quantizations, and known compatibility traps.

Task fit

Use-it-for, skip-it-for, comparisons, and eval-ready shortlists.

Build with the model, then finish the workflow.

Use FixTools for PDFs, JSON, images, metadata, and the developer utilities around your AI stack.

Browse developer tools