AI engineering
Retrieval, evaluation and running language models inside real business systems.
Automate blog and social media posting with Claude, GitHub Actions and Make
An architecture for publishing one researched article a day and turning it into a narrated vertical video for YouTube Shorts, Instagram, Facebook and LinkedIn, with the platform limits that shape it.
Production LLMOps & Enterprise RAG: Low-Latency, Privacy-Preserving Architecture at Scale
A blueprint for building accurate, privacy-preserving Retrieval-Augmented Generation (RAG) systems, covering hybrid dense-sparse search, cross-encoder re-ranking, semantic caching and latency tuning.