My blog

Writing about AI products, backend engineering, and real-world lessons from building and scaling software.

Published on 2026-07-21

Building GapFoundry: An AI Agent That Finds Market Gaps in Minutes

How I built GapFoundry: an AI competitive research platform that automates market gap analysis with cited reports and interactive follow-up. Includes full demo.

GapFoundryAgentic AIAI EngineeringLangGraphLangChainRAGpgvectorFastAPIPythonNext.jsCompetitive ResearchVector DatabasesBackend Engineering
Read post

Published on 2026-07-21

Building Tyvur: An End-to-End AI Interview Platform (Engineering Case Study)

How I built and shipped Tyvur: a live AI interview practice platform with voice, personalized feedback, and a full production stack. Includes a complete product demo.

TyvurVoice AIAI EngineeringLangGraphLangChainFastAPIPythonNext.jsWebSocketsAgentic AISaaSBackend Engineering
Read post

Published on 2026-06-18

Memory Management in LLMs: How AI Actually Remembers Things

Context windows, short-term state, long-term retrieval, RAG, and the memory architectures that separate demo chatbots from production AI systems.

LLMsAI EngineeringGenerative AIMemory SystemsRAGVector DatabasesLangChainLangGraphAI Agents
Read post

Published on 2026-06-10

How LLMs Actually Work, Part 3: Agents and the Future of AI

The final part of a practical deep dive into LLMs - reasoning, mixture of experts, multimodal models, and how LLMs become AI agents that plan, use tools, and execute.

Artificial IntelligenceLarge Language ModelsGenerative AIMachine LearningSoftware EngineeringAI Agents
Read post

Published on 2026-05-29

How LLMs Actually Work, Part 2: Inference, Memory, and RAG

The second part of a practical deep dive into LLMs - parameters, inference, context windows, hallucinations, retrieval augmented generation, and fine-tuning.

Artificial IntelligenceLarge Language ModelsGenerative AIMachine LearningSoftware EngineeringRAG
Read post

Published on 2026-04-30

How LLMs Actually Work, Part 1: Tokens and Transformers

The first part of a practical deep dive into LLMs - from next-token prediction and tokenization through embeddings, attention, transformer blocks, and training.

Artificial IntelligenceLarge Language ModelsGenerative AIMachine LearningSoftware Engineering
Read post

Published on 2026-04-11

Handling Concurrent AI Requests Without Killing Your Server

Learn how to design backend systems that handle multiple AI requests efficiently without crashing, slowing down, or burning unnecessary compute.

Backend EngineeringSystem DesignScalabilityAI SystemsFastAPIPython
Read post

Published on 2026-03-04

Building Tyvur: Lessons From My First Real Product Attempt

What building Tyvur taught me about product reality, validation, and why good engineering alone is not enough.

SaaS JourneyProduct LessonsPythonAgentic AIGenerative AIFastAPIPostgreSQLDockerNext.jsRazorpay
Read post

Published on 2026-02-11

My Passion for Building Software: Why I Keep Showing Up Every Single Day

A personal story of how coding became more than a job, and why I am looking to work with global teams that value ownership, clean systems, and real problem solving.

Software EngineeringDeveloper JourneyBackend DevelopmentGenerative AIFastAPIDjangoRemote WorkProduct ThinkingBuilding in PublicStartups
Read post