Summarizing long emails
Recommended open model
DeepSeek V4 Flash
available in Helmcode
Low-complexity extraction at high volume: near-zero cost per token, with 1M of context and caching.
Open alternative: Gemma 4 12B
open model guide · july 2026
80 enterprise use cases and the open-weight model that solves each one. No closed APIs: everything here can be downloaded, audited and run under your control.
quick selector
Pick the type of task and your main constraint. The recommendation updates instantly; the 80 cases in detail are just below.
What kind of task?
What is your main constraint?
Qwen 3.6APACHE 2.0
The open-source leader in multilingual writing, especially strong in Spanish. Natural drafting, style correction and brand content at the highest open level.
the 80 cases
Filter by area or complexity. Each card gives the primary recommendation, the alternative and why. The green dot marks the models served in Helmcode on a flat rate.
Recommended open model
DeepSeek V4 Flash
available in Helmcode
Low-complexity extraction at high volume: near-zero cost per token, with 1M of context and caching.
Open alternative: Gemma 4 12B
Recommended open model
Qwen 3.6
available in Helmcode
The best open model for natural tone and register in Spanish. When a person signs the email, the writing matters.
Open alternative: GLM-5.2
Recommended open model
Llama 4 Scout
10M of context: whole reports, case files or books, with no chunking and no loss of coherence.
Open alternative: DeepSeek V4 Flash (up to 1M)*
Recommended open model
Whisper large-v3
available in Helmcode
Whisper → V4 Flash pipeline: transcript and minutes 100% on your own infrastructure, with quality very close to the frontier.
Open alternative: DeepSeek V4 Flash (summary)
Recommended open model
DeepSeek V4 Flash
available in Helmcode
The context is already retrieved: extraction plus writing with precise citations. For confidential docs, self-hosting is the only valid option.
Open alternative: Qwen 3.6
Recommended open model
Gemma 4 12B
Classification against clear criteria in milliseconds. It fits on a modest GPU while it processes every incoming email.
Open alternative: DeepSeek V4 Flash
Recommended open model
DeepSeek V4 Flash
available in Helmcode
Entity extraction (what, who, when) with reliable JSON. At this volume, cost is what decides.
Open alternative: Gemma 4 12B
Recommended open model
Qwen3.6-27B
Follow-ups should read as human, not automated. The 27B size holds the quality on a single GPU.
Open alternative: Qwen 3.6 (larger size)
Recommended open model
Qwen 3.6
available in Helmcode
It respects the author’s voice instead of rewriting it: the real difference in editing. The open leader in multilingual style correction.
Open alternative: GLM-5.2
Recommended open model
DeepSeek V4 Flash
available in Helmcode
Frontier-adjacent quality at the lowest cost on the market. For fine cultural adaptation, step up to Qwen 3.6.
Open alternative: Qwen 3.6 (ES/EN/FR/DE/PT)
Recommended open model
qwen3-embedding + rerank
available in Helmcode
The forgotten half of every RAG stack: good embeddings plus reranking decide more than the generator model does.
Open alternative: DeepSeek V4 Flash (synthesis)
Recommended open model
qwen3-embedding + rerank
available in Helmcode
Reordering the retrieved top-k multiplies RAG precision at a minimal marginal cost.
Open alternative: Gemma 4 12B as cross-encoder
Recommended open model
DeepSeek V4 Flash
available in Helmcode
Comprehension plus rephrasing into user language. Ideal for manuals and policies that cannot leave for external APIs.
Open alternative: Qwen3.6-27B
Recommended open model
GLM-5.2
Data synthesis plus narrative: the best open reasoning following complex structured templates.
Open alternative: Qwen 3.6
Recommended open model
GLM-5.2
One conclusion per slide, title as message. Wired to python-pptx or reveal.js: a full data-to-deck pipeline.
Open alternative: Qwen 3.6
Recommended open model
Qwen 3.6
available in Helmcode
Precise language and consistent terminology across long documents. Internal content that should stay at home.
Open alternative: GLM-5.2
Recommended open model
Llama 4 Scout
A 1-3h meeting is 50-150K tokens. Scout handles it whole; a Whisper → Scout pipeline, 100% self-hosted.
Open alternative: DeepSeek V4 Flash*
Recommended open model
DeepSeek V4 Flash
available in Helmcode
RSS/scraping → batch summary and classification → dashboard. Cost per article is practically zero.
Open alternative: Gemma 4 26B
Recommended open model
Gemma 4 12B
One of the simplest tasks on the list: millisecond latency, on-premise, on modest hardware.
Open alternative: DeepSeek V4 Flash
Recommended open model
DeepSeek V4 Flash
available in Helmcode
Task status, alerts and summaries. Milestones and owners of client projects never leave the internal network.
Open alternative: Qwen3.6-27B
Recommended open model
GLM-5.2
The best open reasoning for synthesizing sources you have already gathered. Pair it with your own scraping for the capture step.
Open alternative: DeepSeek V4 Pro
Recommended open model
DeepSeek V4 Pro
The king of bounded generation: 80.6% SWE-bench Verified and 93.5% LiveCodeBench (#1 worldwide). For long-running agents, look at GLM-5.2. K3 is the open ceiling today if you have the cluster to serve it, around 64 GPUs.
Open alternative: Kimi K3, or K2.7 Code on less hardware
Recommended open model
GLM-5.2
The verified open leader in agentic coding: 62.1% SWE-bench Pro and 81.0 on Terminal-Bench 2.1, well ahead of DeepSeek on multi-step work over real repos. K3 is the open ceiling today if you have the cluster to serve it, around 64 GPUs.
Open alternative: Kimi K3, or K2.7 Code on less hardware
Recommended open model
DeepSeek V4 Pro
Spots bugs, debt and security issues by reasoning over the full diff. The code never leaves your infrastructure.
Open alternative: GLM-5.2
Recommended open model
DeepSeek V4 Pro
Edge-case coverage and coherent mocks. High volume: self-hosted, the cost per suite is marginal.
Open alternative: Kimi K2.7 Code
Recommended open model
GLM-5.2
Hours-long multi-file refactors: the area where GLM-5.2 pulls furthest ahead (FrontierSWE, DeepSWE), with 1M of context for the whole repo. K3 is the open ceiling today if you have the cluster to serve it, around 64 GPUs.
Open alternative: Kimi K3, or K2.7 Code on less hardware
Recommended open model
Qwen 3.6
available in Helmcode
Text-to-SQL over your schema with a few-shot prompt. The database schema is sensitive information: better kept at home.
Open alternative: DeepSeek V4 Flash
Recommended open model
Kimi K2.7 Code
Reads the code, writes the docs. K2.7 keeps coherence across modules in large codebases.
Open alternative: GLM-5.2
Recommended open model
DeepSeek V4 Flash
available in Helmcode
Resolves 80% of L1 queries by loading the whole KB into context. Employee and systems data: inside the network.
Open alternative: Gemma 4 26B
Recommended open model
GLM-5.2
Multi-step reasoning connecting unrelated events, with a method (5 Whys, Ishikawa). Confidential logs: self-host.
Open alternative: DeepSeek V4 Pro
Recommended open model
MiniMax M3
Native multimodal plus frontier coding: it reads the screenshot and generates the component. K3 beats it and is multimodal too, but it needs a cluster; M3 gives native vision with a reasonable deployment.
Open alternative: Kimi K3, or K2.7 Code on less hardware
Recommended open model
DeepSeek V4 Flash
available in Helmcode
Training and test datasets at zero marginal cost. The MIT license places no restriction on how you use the outputs.
Open alternative: Qwen 3.6
Recommended open model
Gemma 4 12B
Fine-tuned on your own categories of personal data, it runs at the edge of the pipeline before anything leaves.
Open alternative: DeepSeek V4 Flash
Recommended open model
Whisper large-v3
available in Helmcode
The open standard for multilingual transcription. Customer calls should never pass through an opaque API.
Open alternative: n/a
Recommended open model
Kokoro
available in Helmcode
Natural voice with minimal latency. A full voicebot pipeline: Whisper → LLM → Kokoro, all open.
Open alternative: Whisper (input) + LLM
Recommended open model
MiniMax M3
Processes the document image directly, no separate OCR step. Invoices and IDs demand a 100% internal pipeline.
Open alternative: Gemma 4 26B
Recommended open model
MiniMax M3
Visual QA, catalog tagging and asset verification with the most capable open multimodal model.
Open alternative: Gemma 4 26B
Recommended open model
Gemma 4 26B
available in Helmcode
Text and image in the same call, fine-tunable to your platform’s criteria for more consistency than zero-shot.
Open alternative: MiniMax M3
Recommended open model
Qwen 3.6
available in Helmcode
The best open persuasive writing, with frameworks (SPIN, Challenger) and a customer-centered narrative.
Open alternative: GLM-5.2
Recommended open model
Gemma 4 12B
Fine-tuned on your conversion history, it beats zero-shot by a wide margin: the structural advantage of open source.
Open alternative: DeepSeek V4 Flash
Recommended open model
DeepSeek V4 Flash
available in Helmcode
Briefings from CRM activity. Amounts and clients under negotiation: commercial intelligence that must not leave.
Open alternative: Qwen3.6-27B
Recommended open model
Qwen 3.6
available in Helmcode
Objection handling and personalization by profile. The sales script is confidential strategy: self-host.
Open alternative: GLM-5.2
Recommended open model
GLM-5.2
The most solid open strategic reasoning you can actually deploy. K3 scores higher, but 1.6 TB of weights puts it out of reach unless you already run a cluster.
Open alternative: Kimi K3
Recommended open model
Qwen 3.6
available in Helmcode
Creativity, copywriting and adaptation to brand tone. At high volume, the saving over a closed API is substantial.
Open alternative: GLM-5.2
Recommended open model
Qwen3.6-27B
Dozens of variations per campaign at zero marginal cost, with writing quality that the 27B size covers well.
Open alternative: DeepSeek V4 Flash
Recommended open model
DeepSeek V4 Flash
available in Helmcode
Structured batch generation across the whole catalog or blog. High volume, clear criteria: cost is what rules.
Open alternative: Qwen3.6-27B
Recommended open model
DeepSeek V4 Flash
available in Helmcode
Thousands of opinions in batch with aspect-based analysis. Fine-tuning on your sector’s vocabulary sharpens the result.
Open alternative: Gemma 4 12B (fine-tune)
Recommended open model
Gemma 4 12B
Your platform’s own criteria, fine-tuned, at a practically zero cost per message.
Open alternative: Gemma 4 26B (multimodal)
Recommended open model
Gemma 4 12B
Real-time multi-label, fine-tunable on your ticket history for business-specific precision.
Open alternative: DeepSeek V4 Flash
Recommended open model
DeepSeek V4 Flash
available in Helmcode
A support chatbot with customer data under GDPR: in regulated sectors, self-hosting is not optional.
Open alternative: Qwen3.6-27B
Recommended open model
Qwen 3.6
available in Helmcode
Empathy plus firmness plus a concrete solution. Complaints contain personal data: process it at home.
Open alternative: GLM-5.2
Recommended open model
DeepSeek V4 Flash
available in Helmcode
Thematic classification and sentiment over thousands of open responses, GDPR-compliant by architecture.
Open alternative: Gemma 4 12B
Recommended open model
Qwen3.6-27B
Automatic closure of each ticket with a summary for the CRM. High volume, sound writing, one GPU.
Open alternative: DeepSeek V4 Flash
Recommended open model
Qwen 3.6
available in Helmcode
Telling evidence from generic claims, without bias. Candidate data is GDPR territory: deploy locally.
Open alternative: GLM-5.2
Recommended open model
DeepSeek V4 Flash
available in Helmcode
All the onboarding documentation fits in 1M of context, with no RAG and without internal policies leaving.
Open alternative: Gemma 4 26B
Recommended open model
Qwen 3.6
available in Helmcode
Constructive feedback with nuance. This data can end up in labor proceedings: self-host plus encryption.
Open alternative: GLM-5.2
Recommended open model
DeepSeek V4 Flash
available in Helmcode
HR data is the most sensitive in the company: the conversations do not leave the corporate infrastructure.
Open alternative: Qwen3.6-27B
Recommended open model
Qwen3.6-27B
Persuasive writing with a defined structure, in native Spanish, even for confidential roles.
Open alternative: Qwen 3.6
Recommended open model
Qwen 3.6
available in Helmcode
Instructional design adapted to the learner’s level. In high-volume e-learning, cost per module tends to zero.
Open alternative: GLM-5.2
Recommended open model
Qwen3.6-27B
Good tone and structure on basic hardware. Internal communications stay out of third-party APIs.
Open alternative: Gemma 4 26B
Recommended open model
MiniMax M3
Processes the invoice image directly and returns structured JSON. A 100% internal financial pipeline.
Open alternative: Gemma 4 26B
Recommended open model
DeepSeek V4 Pro
90.1% GPQA and the best open mathematical reasoning. For listed companies or M&A, self-hosting is practically mandatory.
Open alternative: GLM-5.2
Recommended open model
GLM-5.2
Complex patterns over the full transaction history (1M ctx). In banking, the data stays in: the only viable option.
Open alternative: DeepSeek V4 Pro
Recommended open model
DeepSeek V4 Flash
available in Helmcode
High volume of short comparisons: statement lines against ledger entries, with the near-matches flagged for a person. Cost per token is what decides here.
Open alternative: Gemma 4 26B
Recommended open model
DeepSeek V4 Pro
Reads the accounts and writes the reasoning behind a limit, ratio by ratio. That is a solvency judgement on a named company: it belongs self-hosted.
Open alternative: GLM-5.2
Recommended open model
Qwen 3.6
available in Helmcode
The same message has to escalate from a nudge to a formal notice without losing the client. Register weighs more than reasoning, which is where Qwen leads in Spanish.
Open alternative: Gemma 4 26B
Recommended open model
DeepSeek V4 Flash
available in Helmcode
Not extraction, judgement: this dinner is over the per diem, this taxi has no attendees. The policy travels in the prompt and changes without retraining anything.
Open alternative: Gemma 4 12B
Recommended open model
GLM-5.2
The highest open level of nuance comprehension, with MIT license and model card published: auditable end to end.
Open alternative: DeepSeek V4 Pro
Recommended open model
GLM-5.2
With the AI Act penalties applying since Aug 2, an open, traceable stack on EU infrastructure complies by architecture rather than by policy.
Open alternative: Qwen 3.6
Recommended open model
Qwen 3.6
available in Helmcode
Tagging and triage of case files by matter, jurisdiction and urgency, without a single page leaving the firm.
Open alternative: Gemma 4 26B
Recommended open model
Llama 4 Scout
Thousands of documents that only mean something read together: 10M of context takes the whole room in one pass instead of chunking it and losing the cross-references.
Open alternative: GLM-5.2*
Recommended open model
GLM-5.2
Reasoning that has to hold a chain of citations without inventing one. Ground it with RAG on your own database of rulings, never on the model memory.
Open alternative: DeepSeek V4 Pro
Recommended open model
Qwen 3.6
available in Helmcode
Generation rather than analysis: assembling a first draft out of clauses legal has already approved, so the team edits instead of starting from nothing.
Open alternative: GLM-5.2
Recommended open model
GLM-5.2
Watching what changed in a regulation and which internal policies it touches. A different job from checking compliance: this one runs before anybody is out of it.
Open alternative: DeepSeek V4 Flash
Recommended open model
DeepSeek V4 Flash
available in Helmcode
Comparing terms and flagging expirations. Agreed prices are confidential commercial information.
Open alternative: Llama 4 Scout (all contracts)*
Recommended open model
GLM-5.2
Structured reasoning over sources you have already gathered. For web capture, combine it with your own scraping.
Open alternative: DeepSeek V4 Flash
Recommended open model
Qwen 3.6
available in Helmcode
Hundreds of requirements answered one by one, each traceable to a document you already hold. The bid stays confidential until the envelope is opened.
Open alternative: GLM-5.2
Recommended open model
GLM-5.2
The same two hundred questions arrive from every client with the wording changed. Answered from your own evidence base, with a person signing it off.
Open alternative: DeepSeek V4 Flash
Recommended open model
DeepSeek V4 Flash
available in Helmcode
Thousands of short events a day, each needing a route: delay, damage, wrong address or nothing at all. Volume picks the model.
Open alternative: Gemma 4 26B
Recommended open model
MiniMax M3
Reads the picture off the line and writes the defect report against your own criteria. Production images rarely have permission to leave the plant.
Open alternative: Gemma 4 26B
*Llama’s licence is not a free and open-source one (use restrictions and specific conditions for the EU): read it before you build on it.
In production 99.5% of our tokens go through open models, and not on the easy tasks: classifying, extracting, summarizing, drafting, answering and reasoning over documents nobody else has read. What is left is a narrow set of genuinely frontier problems, and the answer to those is to route them to the big model, not to pay frontier prices for the other 99.5%.
go deeper
The 80 cases above roll up into twelve canonical use cases. Each has its own page with architecture, FAQ and deployment in detail.
get started
Pick a case from the list, point the same code at our endpoint and read both outputs side by side. The API is the one you already call, so the test costs you an afternoon. If it does not hold up you have moved nothing, and if it does you already know where the rest of the list goes.
// cookies
We use strictly necessary cookies to run the site and, only with your consent, Google Analytics to understand usage. No advertising, ever — see our Cookie Policy.
// preferences