Skip to content
Houtini.
Contact
Operating · accepting briefs

Make your people and business agent-ready.

AI isn't coming for your job. It's coming to make your job easier, and the exciting bit is that means you, and your team, can concentrate on the creative work, the strategy and the planning. Here's what happens:

  • 01

    The Audit. We find where AI fits in your business.

  • 02

    The Build. We embed in your team and build the bits that take the chores off them, so your people get back to the creative work, the strategy and the planning.

  • 03

    Agent UX. We empower your team to be agent builders, makers, managers and get your website ready for the customers who arrive via AI search or their own AI agents.

Most engagements start with one and grow into another. Plenty of teams just want one. Either is fine.

Founded by
Richard Baxter · founded Builtvisible , exited in 2025 · twenty years in technical SEO and content infrastructure
About Richard →
Worked with
IcelandAirCheapflightsVeryTowergateFairmont HotelsPimaxMoza RacingFanatecColson GroupSim Racing CockpitDriver61
Notes from the work

Read our blog.

From Chat to Deploy: How to Fully Understand Claude, AI Assistants and Claude Code
Article 5 min read

From Chat to Deploy: How to Fully Understand Claude, AI Assistants and Claude Code

I started in the chat box like everyone else. Three years later the boring parts of my work run themselves, from PRD to deploy. This is the ladder, stage by stage, with a starter task for each rung.

How to Do a Technical SEO Audit with Claude
Article 5 min read

How to Do a Technical SEO Audit with Claude

A full technical SEO audit run by conversation in Claude Desktop: Search Console history and a first-party crawl merged into one prioritised, fix-writing report. The setup, the twelve prompts, the priority model, and the honest limits - walked through on a real site.

Tuning vLLM on a Modded 48GB RTX 4090: From 18 to 133 Tokens per Second
Article 5 min read

Tuning vLLM on a Modded 48GB RTX 4090: From 18 to 133 Tokens per Second

Three days of benchmarking vLLM on a modded 48GB RTX 4090: the quantisation trap that costs Ada owners 40% of their speed, the free 1.9x from speculative decoding, the KV compression that never served once, and what cutting 120 watts costs you. Every number measured, every failure kept.

Giving Claude a Local Sidekick: a vLLM Journey Update
Article 5 min read

Giving Claude a Local Sidekick: a vLLM Journey Update

I want Claude to have a powerful local sidekick through houtini-lm. I'm not there yet, but this month I got noticeably closer - a vLLM rebuild, a benching spree on the 48GB 4090, and a face-off that deleted two models in an evening.

Best GPUs for Running Local LLMs (2026): Memory Bandwidth, VRAM and the Cards Worth Buying
Article 5 min read

Best GPUs for Running Local LLMs (2026): Memory Bandwidth, VRAM and the Cards Worth Buying

What to buy in mid-2026 for running local LLMs seriously, from someone who has benched most of it. Why memory bandwidth matters more than FLOPS, why the quantisation format can matter more than the card, and the GPUs worth your money - from the used 3090 floor to the modded 48GB 4090s I ended up buying twice.

Claude Desktop System Requirements: Windows, macOS, Linux (2026)
Article 5 min read

Claude Desktop System Requirements: Windows, macOS, Linux (2026)

What you need to run Claude Desktop in 2026, after Cowork shipped, the Connectors marketplace landed, and the Opus 4.8 / Sonnet 4.6 generation took over. Anthropic's official specs, what real machines need, and where the install falls over.

Browse all articles → Updated weekly · written by Richard
Fancy a chat?

Forty-five minutes. No pitch deck.

Know you need to look at AI in your company but you are not sure what is possible? Jump on a call and we will demonstrate some agentic processes, bring you up to speed with the state of the art, and then discuss your needs.