Built Way.com's API gateway for 5M+ daily requests at 99.95% availability, with Resilience4j, Redis, and Spring Cloud Gateway. At Zafin a Kafka-partitioned delta-processing ETL cut data-load time by 60%.
Vishnu Balachandran
Technical Consultant · Engineering Leader · Technical Architect · Staff Engineer
Technical Consultant and Java technical expert with 8+ years building resilient banking and fintech platforms.
Learning on your own? Browse the courses — check the banner above for an active promo code.
Previously at ZafinWay.comUdaan.comSunTec Business SolutionsFlyTxt Mobile Solutions
My story
Full background →I started as a software engineer at FlyTxt in 2018, chasing down the slow paths in a telecom CRM platform line by line. Eight years later that same instinct, find what is actually broken and fix the real cause, is what let me architect the pricing and billing engine a bank's deal desk runs on every single day.
Most of that time I have built event-driven, Kafka-based systems in Java and Spring Boot that have to be right, not just fast: Zafin's Corporate Pricing Engine prices and bills 2M+ transactions a day at sub-second latency. Way.com's API gateway held up under 5M+ daily requests at 99.95 percent uptime. A profiling pass at FlyTxt cut one platform latency 10x. I have led teams of 10 to 30 engineers through work like this: code review, mentorship, a Spring Boot upgrade across 20+ microservices that cleared 15+ CVEs. An architecture only holds if the team building it does too.
The last couple of years pulled me into AI engineering the same way banking once pulled me into correctness. At Zafin I built an OpenAI-powered Pricing and Billing Copilot: a chatbot with PyTorch-based self-training on the questions people actually ask that cut L1 support resolution time in half. This site runs the same instinct on itself: a RAG pipeline over Hugging Face and NVIDIA models powers the AI Assistant you can talk to right here, grounded in what is actually written on this site, not a guess.
Outside the day job I still build and ship things end to end. This portfolio, its admin backend, its courses, and its chatbot are one project. Temple Accounts and Billing, a small offline billing app I built for temples, now sold here as a downloadable product, is another. If you are trying to get a system like any of these right, that is the kind of problem I like best. Get in touch.
Built Zafins AI-powered Pricing and Billing Copilot: a PyTorch-based chatbot that self-trains on past queries and the answers that resolved them, cutting L1 support resolution time 50%. Also designed this sites own RAG pipeline, embedding its content with NVIDIA and Hugging Face models so its AI Assistant answers grounded in real site content.
Upgraded Spring Boot across 20+ microservices and eliminated 15+ CVEs. At FlyTxt, JProfiler and Hibernate tuning delivered a 10x latency reduction on the serving path.
Zafin's Corporate Pricing Engine prices and bills 2M+ transactions a day at sub-second latency. I was technical SPOC for Deal Manager's first production go-live at ING Bank, at 99.9% uptime on day one.
Mentored 10+ engineers at Zafin; SOLID and CQRS adoption cut defect leakage 40%. At Way.com I led 30+ features and pulled feature time-to-market from 6 weeks to 3.
Automated AWS/EKS deployments at SunTec, cutting manual effort 80%. At Udaan I ran the API gateway on Azure with Resilience4j, dropping P99 latency 35%.
Canary releases, A/B tests, and feature flags at Udaan cut deployment incidents 70% and doubled release velocity. Prometheus, Thanos, and Grafana brought MTTD from 15 minutes to under 3.
PostgreSQL is the system of record behind Zafin's event-driven pricing engine at 2M+ transactions a day. Redis caches Way.com's 5M+ daily-request gateway; Ignite Cache backs SunTec's streaming billing and payments.
Built Zafin's AI Pricing & Billing Copilot, cutting L1 resolution time by 50%. It sits on the live pricing and billing platform, not a side demo.
Shipped 30+ Way.com features and partner integrations — Amazon, getUpside, FuelRewards, Travelers, Cure — adding $1M+ incremental ARR. Client teams consume those through the REST and GraphQL APIs I designed.
From the same work I do for clients
The courses below teach the same patterns I use on real banking and fintech systems, distilled into something you can work through on your own. If your problem is more specific than a course can cover, that's what 1:1 consultancy is for.
Software
All software →Apps and tools I have built and sell directly — download after purchase, no subscription.
Technical writing
Read the blog →Where I work through the architecture decisions behind real production systems — event-driven design, resiliency, and the trade-offs that only show up at scale.
From Embeddings to Answers: Building This Site's Own RAG Pipeline
This site has its own AI Assistant, grounded in its own content through a real retrieval-augmented generation pipeline: embeddings, a relevance gate, access-aware retrieval, and a semantic answer cache. Here is how each piece earns its place.
Building a Self-Training Support Copilot: PyTorch, FastAPI and SSE Streaming on a Java Platform
How we cut L1 billing-support resolution time in half with a Python LLM service that streams answers over SSE and retrains itself on PyTorch from every query it has already resolved, sitting next to a Java and Spring Boot platform it never has to touch directly.
Event-Driven Pricing and Billing at 2M+ Transactions a Day
Architecting a corporate pricing and billing platform where the number has to be reproducible in front of an auditor a year later: why the pricing engine is a Kafka consumer instead of a service you call, how the command and query sides split, and what partitioning actually buys you.
Get new posts & courses by email
No spam, no automated drip — just a note when something new is published.