<?xml version="1.0" encoding="UTF-8"?><rss version="2.0"><channel><title>haimaker.ai Blog</title><description>Technical insights on AI infrastructure, GPU benchmarking, and inference optimization from haimaker.ai</description><link>https://haimaker.ai/</link><item><title>Where to Actually Buy DeepSeek V4: Go vs API vs Gateway</title><link>https://haimaker.ai/blog/deepseek-provider-comparison/</link><guid isPermaLink="true">https://haimaker.ai/blog/deepseek-provider-comparison/</guid><description>DeepSeek repriced V4 on August 16 and every reseller repriced behind it. A buyer&apos;s guide to the three ways to pay for DeepSeek V4 access, and why the provider you pick now changes which checkpoint answers you.</description><pubDate>Thu, 27 Aug 2026 00:00:00 GMT</pubDate></item><item><title>How to Remove Claude Watermarks From Content You Own</title><link>https://haimaker.ai/blog/claude-watermark-removal-guide/</link><guid isPermaLink="true">https://haimaker.ai/blog/claude-watermark-removal-guide/</guid><description>Claude can mark text and files through separate provenance channels. This guide explains a cleanup pipeline for content you own and what the result cannot prove.</description><pubDate>Wed, 12 Aug 2026 00:00:00 GMT</pubDate></item><item><title>Cheapest AI API Keys: Where the Sub-Dollar Models Live</title><link>https://haimaker.ai/blog/cheapest-ai-api-key/</link><guid isPermaLink="true">https://haimaker.ai/blog/cheapest-ai-api-key/</guid><description>What a cheap AI API key actually buys in 2026: the sub-dollar model tier, which free tiers hold up under real load, and why one key beats five.</description><pubDate>Tue, 28 Jul 2026 00:00:00 GMT</pubDate></item><item><title>Best Models for Hermes Agent (July 2026): Tested &amp; Ranked</title><link>https://haimaker.ai/blog/best-models-for-hermes-agent/</link><guid isPermaLink="true">https://haimaker.ai/blog/best-models-for-hermes-agent/</guid><description>Which AI model should you run behind Hermes Agent? Claude Opus 4.8 for coding, Gemini 3.1 Pro for research, MiniMax M3 for around-the-clock budget work. Config values included.</description><pubDate>Thu, 23 Jul 2026 00:00:00 GMT</pubDate></item><item><title>haimaker vs OpenRouter: Which Unified AI API Fits Your Stack?</title><link>https://haimaker.ai/blog/haimaker-vs-openrouter/</link><guid isPermaLink="true">https://haimaker.ai/blog/haimaker-vs-openrouter/</guid><description>A direct comparison of haimaker.ai and OpenRouter: platform fees, provider routing, coding-agent setup, catalogs, and the cases where each one is the right choice.</description><pubDate>Thu, 23 Jul 2026 00:00:00 GMT</pubDate></item><item><title>OpenRouter Alternatives: 8 Unified AI APIs Compared (2026)</title><link>https://haimaker.ai/blog/openrouter-alternatives/</link><guid isPermaLink="true">https://haimaker.ai/blog/openrouter-alternatives/</guid><description>The best OpenRouter alternatives compared: haimaker, LiteLLM, Portkey, Requesty, Together, Fireworks, Cloudflare AI Gateway, and Bifrost. What each one is good at, what OpenRouter&apos;s fees actually cost you, and how to pick.</description><pubDate>Mon, 13 Jul 2026 00:00:00 GMT</pubDate></item><item><title>Which Gemini Model Is Best for Coding? (Free vs Paid)</title><link>https://haimaker.ai/blog/best-gemini-model-for-coding/</link><guid isPermaLink="true">https://haimaker.ai/blog/best-gemini-model-for-coding/</guid><description>Gemini 3 Flash is the best free-tier model for coding, with a generous daily quota and a 1M context window. Here is which Gemini to use for coding, free and paid.</description><pubDate>Tue, 30 Jun 2026 00:00:00 GMT</pubDate></item><item><title>Is Ollama Good for Coding? When Local Wins and When It Doesn&apos;t</title><link>https://haimaker.ai/blog/is-ollama-good-for-coding/</link><guid isPermaLink="true">https://haimaker.ai/blog/is-ollama-good-for-coding/</guid><description>Ollama is genuinely good for coding in 2026, for autocomplete, private edits, and offline work. Here is where local models shine and where agent loops break down.</description><pubDate>Tue, 30 Jun 2026 00:00:00 GMT</pubDate></item><item><title>Claude Opus vs Sonnet vs Haiku: Which Tier to Use (2026)</title><link>https://haimaker.ai/blog/claude-opus-vs-sonnet-vs-haiku/</link><guid isPermaLink="true">https://haimaker.ai/blog/claude-opus-vs-sonnet-vs-haiku/</guid><description>Opus 4.8, Sonnet 4.6, and Haiku 4.5 compared — what each Claude tier is for, when the premium is worth it, and the layered setup most coding teams settle on.</description><pubDate>Thu, 18 Jun 2026 00:00:00 GMT</pubDate></item><item><title>DeepSeek in OpenCode: Setup Guide for V4 Pro &amp; Flash (2026)</title><link>https://haimaker.ai/blog/deepseek-opencode-setup/</link><guid isPermaLink="true">https://haimaker.ai/blog/deepseek-opencode-setup/</guid><description>Wire DeepSeek into OpenCode with a copy-paste provider block. Which V4 model to pick, the credential and config steps, the legacy model-name deadline, and how to fix the usual /models errors.</description><pubDate>Thu, 18 Jun 2026 00:00:00 GMT</pubDate></item><item><title>Best Ollama Coding Models by NVIDIA RTX GPU: VRAM Tier Guide</title><link>https://haimaker.ai/blog/ollama-coding-models-by-nvidia-rtx-vram/</link><guid isPermaLink="true">https://haimaker.ai/blog/ollama-coding-models-by-nvidia-rtx-vram/</guid><description>Which Ollama coding model fits your NVIDIA RTX card, from the 32GB RTX 5090 down to 8GB GPUs. Real VRAM budgets, expected tokens per second, and when to fall back to cloud.</description><pubDate>Mon, 15 Jun 2026 00:00:00 GMT</pubDate></item><item><title>Auto-Router, Improved: Cut Your LLM Bill Without Touching Your Code</title><link>https://haimaker.ai/blog/auto-router-v2-smart-routing/</link><guid isPermaLink="true">https://haimaker.ai/blog/auto-router-v2-smart-routing/</guid><description>The haimaker.ai auto-router now routes every request to the right model, falls back to the cheapest capable one, and learns cost-saving rules from your own traffic. Built for agents and high-volume workloads. One line of code, no quality loss.</description><pubDate>Fri, 12 Jun 2026 00:00:00 GMT</pubDate></item><item><title>How to Run Gemma 4 12B Locally with Ollama and OpenCode</title><link>https://haimaker.ai/blog/gemma-4-12b-ollama-opencode-setup/</link><guid isPermaLink="true">https://haimaker.ai/blog/gemma-4-12b-ollama-opencode-setup/</guid><description>Google&apos;s new Gemma 4 12B is a multimodal open model that runs on a 16GB laptop. Here&apos;s how to pull it with Ollama and wire it into OpenCode as your local coding assistant.</description><pubDate>Thu, 04 Jun 2026 00:00:00 GMT</pubDate></item><item><title>Claude Haiku 4.5 vs Sonnet 4.6: Which to Use for Coding (2026)</title><link>https://haimaker.ai/blog/claude-haiku-4-5-vs-sonnet-4-6/</link><guid isPermaLink="true">https://haimaker.ai/blog/claude-haiku-4-5-vs-sonnet-4-6/</guid><description>Haiku 4.5 and Sonnet 4.6 compared for coding agents — speed, cost, context window, SWE-bench scores, and exactly when each model is the right pick.</description><pubDate>Tue, 02 Jun 2026 00:00:00 GMT</pubDate></item><item><title>How to Use DeepSeek with OpenClaw and OpenCode (2026)</title><link>https://haimaker.ai/blog/deepseek-openclaw-setup/</link><guid isPermaLink="true">https://haimaker.ai/blog/deepseek-openclaw-setup/</guid><description>Wire DeepSeek V4 Pro into OpenClaw or OpenCode. Copy-paste provider config for both agents, which DeepSeek model to pick, and how to dodge the API reliability issues.</description><pubDate>Tue, 02 Jun 2026 00:00:00 GMT</pubDate></item><item><title>Is OpenClaw Free? Costs and Free Setups Explained (2026)</title><link>https://haimaker.ai/blog/is-openclaw-free/</link><guid isPermaLink="true">https://haimaker.ai/blog/is-openclaw-free/</guid><description>OpenClaw the software is 100% free and open source. The cost is the AI model you connect. Here&apos;s what&apos;s free, what isn&apos;t, and how to run it for $0.</description><pubDate>Tue, 02 Jun 2026 00:00:00 GMT</pubDate></item><item><title>Apple Silicon for Local LLMs: When It Pays Off and When It Doesn&apos;t</title><link>https://haimaker.ai/blog/apple-silicon-local-llm-cost-vs-cloud/</link><guid isPermaLink="true">https://haimaker.ai/blog/apple-silicon-local-llm-cost-vs-cloud/</guid><description>An M5 Max can run near-Sonnet models on a laptop. We do the real cost math — hardware depreciation, global electricity prices, and throughput — to show when local inference beats cloud and when it doesn&apos;t.</description><pubDate>Mon, 18 May 2026 00:00:00 GMT</pubDate></item><item><title>Use Codex CLI with Haimaker: Full Setup Guide (2026)</title><link>https://haimaker.ai/blog/how-to-add-haimaker-codex-cli/</link><guid isPermaLink="true">https://haimaker.ai/blog/how-to-add-haimaker-codex-cli/</guid><description>Configure OpenAI&apos;s Codex CLI to route through Haimaker for agentic coding with GPT-5.3-Codex and hundreds of other models. Copy-paste config.toml included.</description><pubDate>Sat, 16 May 2026 00:00:00 GMT</pubDate></item><item><title>How to Use Google Gemini with OpenClaw (API Key Setup, 2026)</title><link>https://haimaker.ai/blog/gemini-api-key-openclaw/</link><guid isPermaLink="true">https://haimaker.ai/blog/gemini-api-key-openclaw/</guid><description>Get a Gemini API key, add Google as a provider in OpenClaw, pick a model, and handle the free-tier limits. Copy-paste config plus a one-key alternative.</description><pubDate>Mon, 11 May 2026 00:00:00 GMT</pubDate></item><item><title>How to Add a Custom Provider to Hermes Agent (2026)</title><link>https://haimaker.ai/blog/hermes-custom-provider-setup/</link><guid isPermaLink="true">https://haimaker.ai/blog/hermes-custom-provider-setup/</guid><description>Connect Hermes Agent to any OpenAI-compatible provider — MiniMax, DeepSeek, Gemini, Grok, OpenAI, or haimaker.ai. Base URLs, API keys, and timeout fixes.</description><pubDate>Mon, 11 May 2026 00:00:00 GMT</pubDate></item></channel></rss>