Production-Grade LLM Orchestration
How I built Gravity OS — a local-first AI operating layer running LLMs offline via Ollama, with agent collaboration and zero cloud dependency.
Deep dives into backend architecture, distributed systems, AI integration, and the hard-won lessons from building production systems at scale.
How I built Gravity OS — a local-first AI operating layer running LLMs offline via Ollama, with agent collaboration and zero cloud dependency.
A deep dive into an event-driven integration layer with Spring Boot and Kafka bridging FLEXCUBE SOAP APIs to modern microservices at 99.9% uptime.
Production patterns for API gateways that handle traffic spikes, cascade failures, and observability — with Node.js and Spring Boot code examples.
A complete guide to an offline AI dev environment — no API keys, no cloud costs, full privacy. Covers model management, quantization, and orchestration.
Event sourcing and CQRS with Kafka and Spring Boot for audit-critical financial workflows: event stores, snapshots, and replay.
Want to discuss any of these topics? I'm always open to technical conversations.
Get in Touch