If you’ve ever wondered how to compare a locally‑hosted LLM with a hosted OpenAI model side‑by‑side, while keeping the whole experiment observable from end‑to‑end, you’re in the right place. This one‑pager walks you through the starter repo that stitches together three distinct paths:
Fabian G. Williams
Principal Product Manager, Microsoft Subscribe to my YouTube.
Recent Posts
Doug Was Right: I Swapped In The MoE, And The Concurrency Math Changed
You Asked, So I Measured: What Concurrency Actually Costs On One Local Model
I Added a Second Local Agent This Week. Here Is the Receipt for Every Human Decision Behind It.
Two Agents, One Local Model: Do They Run in Parallel, or Take Turns? I Measured It.
I Swapped My Local Coding Model Overnight From a Hotel. The Agent Graded the Upgrade Itself.
Categories
- ai21
- building17
- how-to12
- adotob8
- graph5
- mindfullness4
- agentic-commerce3
- ai-agents3
- microsoft3653
- personal-journey3
- ai-development2
- best-practices2
- career2
- engineering2
- gotcha2
- politics2
- product-management2
- wfh2
- agent-architecture1
- agent-engineering1
- ai-commerce1
- azure1
- knowledge-management1
- leadership1
- llama31
- local-llm1
- local-models1
- markdown1
- motorcycle1
- neuroscience1
- nonprofit1
- observability1
- openclaw1
- personal1
- personal-growth1
- semantic-kernel1
- tech-conferences1
- technology-adoption1
- website1
About
Fabian G. Williams aka Fabs Site