GLM 5.2 orquesta, DeepSeek V4 Flash ejecuta
Cómo diseñar sistemas multi-modelo donde un LLM abierto planifica, otro ejecuta y el contexto que se pasan entre ellos es eficiente.
15 artículos
Cómo diseñar sistemas multi-modelo donde un LLM abierto planifica, otro ejecuta y el contexto que se pasan entre ellos es eficiente.
GLM 5.2, el modelo abierto MIT de 744B, ya corre en la infraestructura europea de Helmcode con zero logs y tarifa plana de 150 euros al mes.
La versión tranquila del AI Act: qué obligaciones te aplican de verdad desde el 2 de agosto, cuáles movió el Ómnibus y qué conviene tener listo.
Sube tu documentación una vez o conecta un repositorio de Git, y pregúntale desde el chat. Cada respuesta cita el documento del que sale.
FP8, NVFP4, H200, B200: qué acelera en hardware cada generación de NVIDIA, cuánta VRAM necesita tu modelo y qué GPU comprar. Con datos de producción reales.
La diferencia entre una imagen con 'cara de IA' y una profesional no es el modelo: es el prompt. Guía completa para generar imágenes increíbles en Helmcode.
Nueva incorporación en Helmcode: comunicación, marketing y comunidad. Quién soy, qué vengo a hacer y qué planes tenemos con la comunidad de NaN.
Cómo pasamos Qwen3.6-35B a NVFP4 en una sola RTX PRO 6000, por qué se caía, y el fix de una línea que resultó ser un bug de cuDNN, no de vLLM.
117 billion tokens, 3.68 million requests, 21 countries, and 99.98% uptime. NaN is a community of builders with its own inference infrastructure and a private platform to deploy apps and agents.
In this post we'll learn what parameters and quantization are, so we can figure out how much space AI models take up.
In this post I'll walk you through how the community's inference servers are set up: the hardware we use, the stack we run, and the models we serve.
I've spent several hours over several days documenting and optimizing my entire local environment so I can "mechanize" the work I do every day managing infrastructure for multiple startups.
This post isn't meant to be a guide on how to use Clawd, but rather a look at how we're rolling it out at Helmcode to have an AI Agent that helps us with our day-to-day work managing the Cloud infrastructure of multiple startups.
Kubernetes is one of the most widely used infrastructure tools among companies, and it has become the standard for running containerized applications at scale all over the world.
Before we start, a bit of context. The infrastructure is hosted on AWS and the architecture was based on Serverless services:
// cookies
Solo usamos cookies estrictamente necesarias para que el sitio funcione. Nada de analítica ni publicidad, nunca — consulta la Política de cookies.
// preferencias