How Local LLM Throughput Scales with Multiple Agents
Adding agents does not simply divide a local LLM’s solo token rate between them. Total output can rise even while each agent slows down.
Read articleAleksei Bukhalov
SDET at MariaDB, working on distributed systems, performance, and test infrastructure. Outside work, I experiment with local LLMs and build agent tools.

Adding agents does not simply divide a local LLM’s solo token rate between them. Total output can rise even while each agent slows down.
Read articleAt work
I test how systems behave under load, during failures, and across distributed components. A lot of the job is building the infrastructure needed to make those tests trustworthy.
Correctness, integration, failure modes, and the bugs that only appear with real data and real dependencies.
Load, bottlenecks, race conditions, and behavior that is easy to miss when every component is tested alone.
Environments, CI/CD, observability, and tooling that make difficult failures repeatable and useful.
Selected projects
Experiments with explicit limits, tools for practical work, and infrastructure for testing it.
Small product experiments built to explore narrow ideas and learn the full path from prototype to App Review.
View all experiments