open model guide · july 2026

The right open model for every use case.

80 enterprise use cases and the open-weight model that solves each one. No closed APIs: everything here can be downloaded, audited and run under your control.

quick selector

Two questions, one model.

Pick the type of task and your main constraint. The recommendation updates instantly; the 80 cases in detail are just below.

What kind of task?

What is your main constraint?

Qwen 3.6APACHE 2.0

The open-source leader in multilingual writing, especially strong in Spanish. Natural drafting, style correction and brand content at the highest open level.

Apache 2.0256K ctx ● available in Helmcode

see_the_cases_for_this_model →

the 80 cases

Every use case, its open model.

Filter by area or complexity. Each card gives the primary recommendation, the alternative and why. The green dot marks the models served in Helmcode on a flat rate.

area
complexity
model: 80 / 80 cases
01
Universal Low

Summarizing long emails

Recommended open model

DeepSeek V4 Flash

available in Helmcode

MIT · 1M ctx

Low-complexity extraction at high volume: near-zero cost per token, with 1M of context and caching.

Open alternative: Gemma 4 12B

02
Universal Medium

Writing professional emails

Recommended open model

Qwen 3.6

available in Helmcode

Apache 2.0 · 256K ctx

The best open model for natural tone and register in Spanish. When a person signs the email, the writing matters.

Open alternative: GLM-5.2

03
Universal Low

Summarizing very long documents

Recommended open model

Llama 4 Scout

Llama · 10M ctx

10M of context: whole reports, case files or books, with no chunking and no loss of coherence.

Open alternative: DeepSeek V4 Flash (up to 1M)*

04
Universal Medium

Transcribing and summarizing meetings

Recommended open model

Whisper large-v3

available in Helmcode

MIT · multilingual STT

Whisper → V4 Flash pipeline: transcript and minutes 100% on your own infrastructure, with quality very close to the frontier.

Open alternative: DeepSeek V4 Flash (summary)

05
Universal Low

RAG over internal documentation

Recommended open model

DeepSeek V4 Flash

available in Helmcode

MIT · 1M ctx

The context is already retrieved: extraction plus writing with precise citations. For confidential docs, self-hosting is the only valid option.

Open alternative: Qwen 3.6

06
Universal Low

Prioritizing the inbox

Recommended open model

Gemma 4 12B

Gemma · 128K ctx · 1 GPU

Classification against clear criteria in milliseconds. It fits on a modest GPU while it processes every incoming email.

Open alternative: DeepSeek V4 Flash

07
Universal Low

Creating tasks from emails and chats

Recommended open model

DeepSeek V4 Flash

available in Helmcode

MIT · 1M ctx

Entity extraction (what, who, when) with reliable JSON. At this volume, cost is what decides.

Open alternative: Gemma 4 12B

08
Universal Low

Automated email follow-ups

Recommended open model

Qwen3.6-27B

Apache 2.0 · 1 GPU 24GB

Follow-ups should read as human, not automated. The 27B size holds the quality on a single GPU.

Open alternative: Qwen 3.6 (larger size)

09
Universal Medium

Proofreading and improving text

Recommended open model

Qwen 3.6

available in Helmcode

Apache 2.0 · 256K ctx

It respects the author’s voice instead of rewriting it: the real difference in editing. The open leader in multilingual style correction.

Open alternative: GLM-5.2

10
Universal Low

Bulk translation

Recommended open model

DeepSeek V4 Flash

available in Helmcode

MIT · 1M ctx

Frontier-adjacent quality at the lowest cost on the market. For fine cultural adaptation, step up to Qwen 3.6.

Open alternative: Qwen 3.6 (ES/EN/FR/DE/PT)

11
Universal Low

Semantic search over the KB

Recommended open model

qwen3-embedding + rerank

available in Helmcode

Apache 2.0 · embeddings

The forgotten half of every RAG stack: good embeddings plus reranking decide more than the generator model does.

Open alternative: DeepSeek V4 Flash (synthesis)

12
Universal Low

Reranking search results

Recommended open model

qwen3-embedding + rerank

available in Helmcode

Apache 2.0 · embeddings

Reordering the retrieved top-k multiplies RAG precision at a minimal marginal cost.

Open alternative: Gemma 4 12B as cross-encoder

13
Universal Low

Generating FAQs from documentation

Recommended open model

DeepSeek V4 Flash

available in Helmcode

MIT · 1M ctx

Comprehension plus rephrasing into user language. Ideal for manuals and policies that cannot leave for external APIs.

Open alternative: Qwen3.6-27B

14
Universal Medium

Executive reports

Recommended open model

GLM-5.2

MIT · 1M ctx · 744B

Data synthesis plus narrative: the best open reasoning following complex structured templates.

Open alternative: Qwen 3.6

15
Universal Medium

Presentations from data

Recommended open model

GLM-5.2

MIT · 1M ctx · 744B

One conclusion per slide, title as message. Wired to python-pptx or reveal.js: a full data-to-deck pipeline.

Open alternative: Qwen 3.6

16
Universal Medium

Policies and procedures

Recommended open model

Qwen 3.6

available in Helmcode

Apache 2.0 · 256K ctx

Precise language and consistent terminology across long documents. Internal content that should stay at home.

Open alternative: GLM-5.2

17
Universal Medium

Meeting agent (agenda + minutes)

Recommended open model

Llama 4 Scout

Llama · 10M ctx

A 1-3h meeting is 50-150K tokens. Scout handles it whole; a Whisper → Scout pipeline, 100% self-hosted.

Open alternative: DeepSeek V4 Flash*

18
Universal Low

News and intelligence summaries

Recommended open model

DeepSeek V4 Flash

available in Helmcode

MIT · 1M ctx

RSS/scraping → batch summary and classification → dashboard. Cost per article is practically zero.

Open alternative: Gemma 4 26B

19
Universal Low

Detecting urgency in messages

Recommended open model

Gemma 4 12B

Gemma · 128K ctx · 1 GPU

One of the simplest tasks on the list: millisecond latency, on-premise, on modest hardware.

Open alternative: DeepSeek V4 Flash

20
Universal Low

Project pipeline management

Recommended open model

DeepSeek V4 Flash

available in Helmcode

MIT · 1M ctx

Task status, alerts and summaries. Milestones and owners of client projects never leave the internal network.

Open alternative: Qwen3.6-27B

21
Universal Medium

Research over your own documents

Recommended open model

GLM-5.2

MIT · 1M ctx · 744B

The best open reasoning for synthesizing sources you have already gathered. Pair it with your own scraping for the capture step.

Open alternative: DeepSeek V4 Pro

22
Code & IT High

Code generation

Recommended open model

DeepSeek V4 Pro

MIT · 1M ctx · 1.6T/49B act

The king of bounded generation: 80.6% SWE-bench Verified and 93.5% LiveCodeBench (#1 worldwide). For long-running agents, look at GLM-5.2. K3 is the open ceiling today if you have the cluster to serve it, around 64 GPUs.

Open alternative: Kimi K3, or K2.7 Code on less hardware

23
Code & IT High

Autonomous coding agents

Recommended open model

GLM-5.2

MIT · 1M ctx · 744B

The verified open leader in agentic coding: 62.1% SWE-bench Pro and 81.0 on Terminal-Bench 2.1, well ahead of DeepSeek on multi-step work over real repos. K3 is the open ceiling today if you have the cluster to serve it, around 64 GPUs.

Open alternative: Kimi K3, or K2.7 Code on less hardware

24
Code & IT Medium

Automated code review

Recommended open model

DeepSeek V4 Pro

MIT · 1M ctx · 1.6T/49B act

Spots bugs, debt and security issues by reasoning over the full diff. The code never leaves your infrastructure.

Open alternative: GLM-5.2

25
Code & IT Medium

Test generation

Recommended open model

DeepSeek V4 Pro

MIT · 1M ctx · 1.6T/49B act

Edge-case coverage and coherent mocks. High volume: self-hosted, the cost per suite is marginal.

Open alternative: Kimi K2.7 Code

26
Code & IT High

Long migrations and refactors

Recommended open model

GLM-5.2

MIT · 1M ctx · 744B

Hours-long multi-file refactors: the area where GLM-5.2 pulls furthest ahead (FrontierSWE, DeepSWE), with 1M of context for the whole repo. K3 is the open ceiling today if you have the cluster to serve it, around 64 GPUs.

Open alternative: Kimi K3, or K2.7 Code on less hardware

27
Code & IT Low

SQL from natural language

Recommended open model

Qwen 3.6

available in Helmcode

Apache 2.0 · 256K ctx

Text-to-SQL over your schema with a few-shot prompt. The database schema is sensitive information: better kept at home.

Open alternative: DeepSeek V4 Flash

28
Code & IT Medium

Automated technical documentation

Recommended open model

Kimi K2.7 Code

Open · 256K ctx · agentic

Reads the code, writes the docs. K2.7 keeps coherence across modules in large codebases.

Open alternative: GLM-5.2

29
Code & IT Low

L1 support / IT helpdesk agent

Recommended open model

DeepSeek V4 Flash

available in Helmcode

MIT · 1M ctx

Resolves 80% of L1 queries by loading the whole KB into context. Employee and systems data: inside the network.

Open alternative: Gemma 4 26B

30
Code & IT High

Log analysis and root cause (RCA)

Recommended open model

GLM-5.2

MIT · 1M ctx · 744B

Multi-step reasoning connecting unrelated events, with a method (5 Whys, Ishikawa). Confidential logs: self-host.

Open alternative: DeepSeek V4 Pro

31
Code & IT Medium

Frontend and UI from mockups

Recommended open model

MiniMax M3

Open · 1M ctx · multimodal

Native multimodal plus frontier coding: it reads the screenshot and generates the component. K3 beats it and is multimodal too, but it needs a cluster; M3 gives native vision with a reasonable deployment.

Open alternative: Kimi K3, or K2.7 Code on less hardware

32
Code & IT Low

Synthetic data generation

Recommended open model

DeepSeek V4 Flash

available in Helmcode

MIT · 1M ctx

Training and test datasets at zero marginal cost. The MIT license places no restriction on how you use the outputs.

Open alternative: Qwen 3.6

33
Code & IT Low

PII detection and anonymization

Recommended open model

Gemma 4 12B

Gemma · 128K ctx · 1 GPU

Fine-tuned on your own categories of personal data, it runs at the edge of the pipeline before anything leaves.

Open alternative: DeepSeek V4 Flash

34
Voice & Multimodal Low

Call transcription (STT)

Recommended open model

Whisper large-v3

available in Helmcode

MIT · multilingual STT

The open standard for multilingual transcription. Customer calls should never pass through an opaque API.

Open alternative: n/a

35
Voice & Multimodal Low

Voicebots and speech synthesis (TTS)

Recommended open model

Kokoro

available in Helmcode

Apache 2.0 · TTS

Natural voice with minimal latency. A full voicebot pipeline: Whisper → LLM → Kokoro, all open.

Open alternative: Whisper (input) + LLM

36
Voice & Multimodal Low

OCR + extraction from scans

Recommended open model

MiniMax M3

Open · 1M ctx · multimodal

Processes the document image directly, no separate OCR step. Invoices and IDs demand a 100% internal pipeline.

Open alternative: Gemma 4 26B

37
Voice & Multimodal Medium

Product image analysis

Recommended open model

MiniMax M3

Open · 1M ctx · multimodal

Visual QA, catalog tagging and asset verification with the most capable open multimodal model.

Open alternative: Gemma 4 26B

38
Voice & Multimodal Low

Multimodal moderation

Recommended open model

Gemma 4 26B

available in Helmcode

Gemma · 256K ctx · multimodal

Text and image in the same call, fine-tunable to your platform’s criteria for more consistency than zero-shot.

Open alternative: MiniMax M3

39
Sales High

Sales proposals from a briefing

Recommended open model

Qwen 3.6

available in Helmcode

Apache 2.0 · 256K ctx

The best open persuasive writing, with frameworks (SPIN, Challenger) and a customer-centered narrative.

Open alternative: GLM-5.2

40
Sales Low

Lead scoring and prioritization

Recommended open model

Gemma 4 12B

Gemma · 128K ctx · 1 GPU

Fine-tuned on your conversion history, it beats zero-shot by a wide margin: the structural advantage of open source.

Open alternative: DeepSeek V4 Flash

41
Sales Low

Opportunity tracking

Recommended open model

DeepSeek V4 Flash

available in Helmcode

MIT · 1M ctx

Briefings from CRM activity. Amounts and clients under negotiation: commercial intelligence that must not leave.

Open alternative: Qwen3.6-27B

42
Sales Medium

Sales scripts and calls

Recommended open model

Qwen 3.6

available in Helmcode

Apache 2.0 · 256K ctx

Objection handling and personalization by profile. The sales script is confidential strategy: self-host.

Open alternative: GLM-5.2

43
Sales High

Negotiation preparation

Recommended open model

GLM-5.2

MIT · 1M ctx · 744B

The most solid open strategic reasoning you can actually deploy. K3 scores higher, but 1.6 TB of weights puts it out of reach unless you already run a cluster.

Open alternative: Kimi K3

44
Marketing Medium

Marketing content

Recommended open model

Qwen 3.6

available in Helmcode

Apache 2.0 · 256K ctx

Creativity, copywriting and adaptation to brand tone. At high volume, the saving over a closed API is substantial.

Open alternative: GLM-5.2

45
Marketing Low

Copy and A/B variations

Recommended open model

Qwen3.6-27B

Apache 2.0 · 1 GPU 24GB

Dozens of variations per campaign at zero marginal cost, with writing quality that the 27B size covers well.

Open alternative: DeepSeek V4 Flash

46
Marketing Low

SEO: briefs and meta descriptions

Recommended open model

DeepSeek V4 Flash

available in Helmcode

MIT · 1M ctx

Structured batch generation across the whole catalog or blog. High volume, clear criteria: cost is what rules.

Open alternative: Qwen3.6-27B

47
Marketing Low

Review sentiment analysis

Recommended open model

DeepSeek V4 Flash

available in Helmcode

MIT · 1M ctx

Thousands of opinions in batch with aspect-based analysis. Fine-tuning on your sector’s vocabulary sharpens the result.

Open alternative: Gemma 4 12B (fine-tune)

48
Marketing Low

Content moderation

Recommended open model

Gemma 4 12B

Gemma · 128K ctx · 1 GPU

Your platform’s own criteria, fine-tuned, at a practically zero cost per message.

Open alternative: Gemma 4 26B (multimodal)

49
Customer support Low

Ticket classification and routing

Recommended open model

Gemma 4 12B

Gemma · 128K ctx · 1 GPU

Real-time multi-label, fine-tunable on your ticket history for business-specific precision.

Open alternative: DeepSeek V4 Flash

50
Customer support Low

Automated customer replies (FAQ)

Recommended open model

DeepSeek V4 Flash

available in Helmcode

MIT · 1M ctx

A support chatbot with customer data under GDPR: in regulated sectors, self-hosting is not optional.

Open alternative: Qwen3.6-27B

51
Customer support Medium

Handling complaints

Recommended open model

Qwen 3.6

available in Helmcode

Apache 2.0 · 256K ctx

Empathy plus firmness plus a concrete solution. Complaints contain personal data: process it at home.

Open alternative: GLM-5.2

52
Customer support Low

Survey and NPS analysis

Recommended open model

DeepSeek V4 Flash

available in Helmcode

MIT · 1M ctx

Thematic classification and sentiment over thousands of open responses, GDPR-compliant by architecture.

Open alternative: Gemma 4 12B

53
Customer support Low

Support conversation summaries

Recommended open model

Qwen3.6-27B

Apache 2.0 · 1 GPU 24GB

Automatic closure of each ticket with a summary for the CRM. High volume, sound writing, one GPU.

Open alternative: DeepSeek V4 Flash

54
HR Medium

CV and candidate screening

Recommended open model

Qwen 3.6

available in Helmcode

Apache 2.0 · 256K ctx

Telling evidence from generic claims, without bias. Candidate data is GDPR territory: deploy locally.

Open alternative: GLM-5.2

55
HR Low

New employee onboarding

Recommended open model

DeepSeek V4 Flash

available in Helmcode

MIT · 1M ctx

All the onboarding documentation fits in 1M of context, with no RAG and without internal policies leaving.

Open alternative: Gemma 4 26B

56
HR High

Performance reviews

Recommended open model

Qwen 3.6

available in Helmcode

Apache 2.0 · 256K ctx

Constructive feedback with nuance. This data can end up in labor proceedings: self-host plus encryption.

Open alternative: GLM-5.2

57
HR Low

HR chatbot for employees

Recommended open model

DeepSeek V4 Flash

available in Helmcode

MIT · 1M ctx

HR data is the most sensitive in the company: the conversations do not leave the corporate infrastructure.

Open alternative: Qwen3.6-27B

58
HR Low

Job postings

Recommended open model

Qwen3.6-27B

Apache 2.0 · 1 GPU 24GB

Persuasive writing with a defined structure, in native Spanish, even for confidential roles.

Open alternative: Qwen 3.6

59
HR Medium

Training and educational materials

Recommended open model

Qwen 3.6

available in Helmcode

Apache 2.0 · 256K ctx

Instructional design adapted to the learner’s level. In high-volume e-learning, cost per module tends to zero.

Open alternative: GLM-5.2

60
HR Low

Internal comms and newsletters

Recommended open model

Qwen3.6-27B

Apache 2.0 · 1 GPU 24GB

Good tone and structure on basic hardware. Internal communications stay out of third-party APIs.

Open alternative: Gemma 4 26B

61
Finance Low

Invoices and expenses (extraction)

Recommended open model

MiniMax M3

Open · 1M ctx · multimodal

Processes the invoice image directly and returns structured JSON. A 100% internal financial pipeline.

Open alternative: Gemma 4 26B

62
Finance High

Financial analysis and reporting

Recommended open model

DeepSeek V4 Pro

MIT · 1M ctx · 1.6T/49B act

90.1% GPQA and the best open mathematical reasoning. For listed companies or M&A, self-hosting is practically mandatory.

Open alternative: GLM-5.2

63
Finance High

Risk and fraud detection

Recommended open model

GLM-5.2

MIT · 1M ctx · 744B

Complex patterns over the full transaction history (1M ctx). In banking, the data stays in: the only viable option.

Open alternative: DeepSeek V4 Pro

64
Finance Low

Bank reconciliation and matching

Recommended open model

DeepSeek V4 Flash

available in Helmcode

MIT · 1M ctx

High volume of short comparisons: statement lines against ledger entries, with the near-matches flagged for a person. Cost per token is what decides here.

Open alternative: Gemma 4 26B

65
Finance High

Credit analysis and underwriting

Recommended open model

DeepSeek V4 Pro

MIT · 1M ctx · 1.6T/49B act

Reads the accounts and writes the reasoning behind a limit, ratio by ratio. That is a solvency judgement on a named company: it belongs self-hosted.

Open alternative: GLM-5.2

66
Finance Medium

Collections and payment reminders

Recommended open model

Qwen 3.6

available in Helmcode

Apache 2.0 · 256K ctx

The same message has to escalate from a nudge to a formal notice without losing the client. Register weighs more than reasoning, which is where Qwen leads in Spanish.

Open alternative: Gemma 4 26B

67
Finance Low

Expense policy checks

Recommended open model

DeepSeek V4 Flash

available in Helmcode

MIT · 1M ctx

Not extraction, judgement: this dinner is over the per diem, this taxi has no attendees. The policy travels in the prompt and changes without retraining anything.

Open alternative: Gemma 4 12B

68
Legal High

Contract and clause analysis

Recommended open model

GLM-5.2

MIT · 1M ctx · 744B

The highest open level of nuance comprehension, with MIT license and model card published: auditable end to end.

Open alternative: DeepSeek V4 Pro

69
Legal High

Compliance and regulation

Recommended open model

GLM-5.2

MIT · 1M ctx · 744B

With the AI Act penalties applying since Aug 2, an open, traceable stack on EU infrastructure complies by architecture rather than by policy.

Open alternative: Qwen 3.6

70
Legal Low

Legal document classification

Recommended open model

Qwen 3.6

available in Helmcode

Apache 2.0 · 256K ctx

Tagging and triage of case files by matter, jurisdiction and urgency, without a single page leaving the firm.

Open alternative: Gemma 4 26B

71
Legal High

Due diligence in a data room

Recommended open model

Llama 4 Scout

Llama · 10M ctx

Thousands of documents that only mean something read together: 10M of context takes the whole room in one pass instead of chunking it and losing the cross-references.

Open alternative: GLM-5.2*

72
Legal High

Case law and precedent research

Recommended open model

GLM-5.2

MIT · 1M ctx · 744B

Reasoning that has to hold a chain of citations without inventing one. Ground it with RAG on your own database of rulings, never on the model memory.

Open alternative: DeepSeek V4 Pro

73
Legal Medium

Drafting from a clause library

Recommended open model

Qwen 3.6

available in Helmcode

Apache 2.0 · 256K ctx

Generation rather than analysis: assembling a first draft out of clauses legal has already approved, so the team edits instead of starting from nothing.

Open alternative: GLM-5.2

74
Legal Medium

Regulatory change monitoring

Recommended open model

GLM-5.2

MIT · 1M ctx · 744B

Watching what changed in a regulation and which internal policies it touches. A different job from checking compliance: this one runs before anybody is out of it.

Open alternative: DeepSeek V4 Flash

75
Operations & Strategy Low

Supplier and procurement management

Recommended open model

DeepSeek V4 Flash

available in Helmcode

MIT · 1M ctx

Comparing terms and flagging expirations. Agreed prices are confidential commercial information.

Open alternative: Llama 4 Scout (all contracts)*

76
Operations & Strategy Medium

Competitor and market analysis

Recommended open model

GLM-5.2

MIT · 1M ctx · 744B

Structured reasoning over sources you have already gathered. For web capture, combine it with your own scraping.

Open alternative: DeepSeek V4 Flash

77
Operations & Strategy Medium

RFP and tender responses

Recommended open model

Qwen 3.6

available in Helmcode

Apache 2.0 · 256K ctx

Hundreds of requirements answered one by one, each traceable to a document you already hold. The bid stays confidential until the envelope is opened.

Open alternative: GLM-5.2

78
Operations & Strategy Medium

Security and vendor risk questionnaires

Recommended open model

GLM-5.2

MIT · 1M ctx · 744B

The same two hundred questions arrive from every client with the wording changed. Answered from your own evidence base, with a person signing it off.

Open alternative: DeepSeek V4 Flash

79
Operations & Strategy Low

Logistics exception triage

Recommended open model

DeepSeek V4 Flash

available in Helmcode

MIT · 1M ctx

Thousands of short events a day, each needing a route: delay, damage, wrong address or nothing at all. Volume picks the model.

Open alternative: Gemma 4 26B

80
Operations & Strategy Medium

Quality inspection from photos

Recommended open model

MiniMax M3

Open · 1M ctx · multimodal

Reads the picture off the line and writes the defect report against your own criteria. Production images rarely have permission to leave the plant.

Open alternative: Gemma 4 26B

*Llama’s licence is not a free and open-source one (use restrictions and specific conditions for the EU): read it before you build on it.

available in Helmcode (flat rate, EU, zero logs) complexity = how demanding the task is, not the deployment Gemma and Llama ship under their own terms, with use restrictions: check them per case

The gap with closed models only exists at the frontier edge. Almost no real work lives there.

In production 99.5% of our tokens go through open models, and not on the easy tasks: classifying, extracting, summarizing, drafting, answering and reasoning over documents nobody else has read. What is left is a narrow set of genuinely frontier problems, and the answer to those is to route them to the big model, not to pay frontier prices for the other 99.5%.

get started

Start with one case, not with a migration.

Pick a case from the list, point the same code at our endpoint and read both outputs side by side. The API is the one you already call, so the test costs you an afternoon. If it does not hold up you have moved nothing, and if it does you already know where the rest of the list goes.