Thinking about the hard parts.

Articles on the problems that determine whether AI projects ship or stall written from twenty years of building systems for domains that resist simple answers.

Articles

Reliability

Building reliable systems out of an unreliable ingredient

LLMs will produce confident wrong answers. The question isn't whether AI makes mistakes it's whether you can build reliable systems using this unpredictable ingredient. The answer is yes, and the method is evaluation decomposition.

Read the article

More articles in progress.

Start here

What are you trying to build?

If you're evaluating where AI fits in your operation or you have a project stalled between demo and production, that's the conversation Skowak is built for.

Start a conversation See how we work