Fractional AI engineering.

Lights Out Labs works on retainer with solo founders and small teams. We build agent workflows, production infrastructure, and business automation, and we leave you with systems you can run yourself.

Agent workflows Production infrastructure Business automation Lights out
What small teams run into
Agents

Your agents ship faster than you can review.

We set up merge gates, review automation, and PR-stack workflows so the queue keeps moving without a review team.

Infrastructure

Production has no platform team to page.

We build Kubernetes clusters, self-hosted models, and high-availability setups that one person can operate.

Direction

Every week ships a new model and a new tool.

We review your stack and recommend specific tools and models you can commit to.

Operations

Founder hours go to work a machine should do.

We automate billing, reporting, support, and the glue between the tools you already use.

From the blog All posts →

Qwen3.8 hits 81.4 tokens/s on an RTX 3090

Local Qwen3.8-27B peaks at 81.4 tokens/s with a few perf adjustments.

Muse Glimmer is a flop

Muse Glimmer is a flop: chatty, prone to hallucinations, bad at tool calling, slower than Qwen and only half of Qwen 3.6's context.

Tell us what is slowing you down.

Book a thirty-minute call or write to us.

Book an intro call

Write to Lights Out Labs

Send a note and we will get back to you.