I ship 0-to-1 systems, then go a layer below them.
A real-estate platform I built and still work on serves ~30K daily visitors.
I architected an offline-first POS that keeps 11 terminals selling through
internet outages. And I traced multi-minute reports in a 15-year-old database
down to missing keys and absent indexes — got them under 3 seconds.
Right now I'm going deep on LLM inference performance: what a token actually
costs, where the latency goes, and when an optimisation stops paying for
itself. I publish every measurement, including the ones that didn't work:
<a href="https://imvivekvermaa.in/log" rel="nofollow">https://imvivekvermaa.in/log
imvivekvermaaa · · focus · HN ↗
Remote: Yes (worldwide)
Willing to relocate: Yes
Technologies: TypeScript, React, Next.js, Node.js, Python, PostgreSQL, AWS, React Native/Expo, LLM apps (RAG, structured extraction, agents), vLLM
Résumé: <a href="https://imvivekvermaa.in/vivek-verma-resume.pdf" rel="nofollow">https://imvivekvermaa.in/vivek-verma-resume.pdf
GitHub: <a href="https://github.com/imvivekvermaa" rel="nofollow">https://github.com/imvivekvermaa
Website: <a href="https://imvivekvermaa.in" rel="nofollow">https://imvivekvermaa.in
Email: imvivekvermaa@gmail.com
I ship 0-to-1 systems, then go a layer below them.
A real-estate platform I built and still work on serves ~30K daily visitors. I architected an offline-first POS that keeps 11 terminals selling through internet outages. And I traced multi-minute reports in a 15-year-old database down to missing keys and absent indexes — got them under 3 seconds.
Right now I'm going deep on LLM inference performance: what a token actually costs, where the latency goes, and when an optimisation stops paying for itself. I publish every measurement, including the ones that didn't work: <a href="https://imvivekvermaa.in/log" rel="nofollow">https://imvivekvermaa.in/log