Tutorial posts from the AIHubMix team.
Migrating from Claude Haiku 4.5 to 5.5: Five 400 Errors and the Quiet ChangesChanging claude-haiku-4-5 to claude-haiku-5-5 is the smallest part of this migration. Five request patterns that worked on Haiku 4.5 now return a 400 error, and several more changes fail no request but alter what you get back, what it costs, or how the model behaves inside an agent. Anthropic says existing Haiku 4.5 prompts should work well on Haiku 5.5 without changes. The request code around those prompts is a different story. This post lists each problem as you'll meet it: what you'll see,
Claude Haiku 5.5 Effort Levels: Medium Is the Default, Low Is Often EnoughClaude Haiku 5.5 is the first Haiku with an effort setting, and the setting moves the bill more than anything else you control. In Artificial Analysis's independent runs, Haiku 5.5 at max effort scored 43 on its Intelligence Index and at low effort scored 29. Max also used 440 million output tokens to get through the index; low used 32 million. So the short answer: leave most work at the default, medium. Drop high-volume, simple routes to low. Raise knowledge work and strict instruction follow
Page Assist in Practice: Four Everyday Browser Workflows with Local and Cloud ModelsMost AI chat tools live in their own tab. You copy text out of a page, paste it into a chat window, and copy the answer back. Page Assist removes that round trip: it is an open-source browser extension that opens a model in a sidebar next to whatever page you are reading, plus a full-tab web UI for longer conversations. It started as a front end for local models running in Ollama, and that is still the default. But it also accepts any OpenAI-compatible endpoint, which means the same sidebar can
Migrating to GPT-6.1 Sol: 9 Things That Can Go WrongMoving from GPT-6 Sol to 6.1 Sol looks like a one-line change, and the price is the same. But there are a few breaking changes and some shifts in behavior, so changing only the model name can get you errors, a surprise bill, or an agent that behaves differently. OpenAI's GPT-6 migration guide covers most of the official changes. This post adds the things that tend to bite in practice. They're ordered from "fails loudly" to "fails quietly." 1. reasoning_effort: "none" returns a 400 What you'l
Choosing a Reasoning Effort for GPT-6.1 Sol: low to maxGPT-6.1 Sol has five reasoning effort levels: low, medium (the default), high, xhigh, and max. The big change from GPT-6 Sol is that none and minimal are gone, so low is now the floor. This one setting drives both latency and cost. You never see reasoning tokens, but you pay for them at the output rate ($10 per million for 6.1 Sol on AIHubMix), and they take up room in the context window. Pick the wrong level and you can easily pay several times more than you need to. What each level is for
How to Use a Real Face in Seedance 2.5 with AIHubMix Real-Human AssetsAI video becomes much more useful when the person on screen stays recognizable from one scene to the next. AIHubMix real-human assets provide a consent-based way to use an approved face in Seedance video generation while keeping authorization, asset processing, and generation as separate, verifiable steps. This guide explains what real-human video is, how the AIHubMix workflow works, why it is useful, and how to create a Seedance 2.5 video from a real face. It also explains the difference betwe
GLM-5.3 Hands-on Guide: Always-on Thinking, Three Effort Levels, and the API Support MatrixAn August 2026 guide to calling GLM-5.3: always-on thinking with three reasoning_effort levels, reasoning summaries, parallel tool calls, structured output, and automatic caching — with tested examples for the AIHubMix Chat, Responses, and Messages APIs.
DeepSeek V4 Pro (0813): Thinking Passback & 3-API MatrixDeepSeek V4 Pro (0813) hands-on guide: thinking toggle and reasoning_effort levels, mandatory thinking-history passback, tools, caching, and a 3-API matrix.
Kimi K3 Hands-On Guide: New Parameters & API Support MatrixJuly 2026 Kimi K3 guide: reasoning_effort max, thinking history, dynamic tool loading, structured output, auto caching, partial prefix, and vision inputs.
Claude Opus 4.7 New Parameters GuideThis article covers two key changes to reasoning control in Claude Opus 4.7, along with complete usage instructions for both the AIHubmix native API and the Chat unified interface. See also: Anthropic official announcement and model change log.