Now
What I'm doing now
A running snapshot of what I'm building, exploring, and reading — updated by hand, not scraped. Inspired by Derek Sivers' /now page.
Last updated Jul 2026
Building now
- DBWhisper evals — the natural-language-to-SQL agent is now measured on execution accuracy, not vibes: a golden-query harness (82% exact, 100% fail-closed) plus a scoped Spider dev run (73%, 101/139). Numbers and method live on Evals.
- This site as a product — turning the portfolio into a running notebook: Notes, per-product changelogs, and the grounded "Ask Friday" assistant, all on the same $0 stack.
- Keeping four products live — DBWhisper, TradePulse, CrownWager, and LLM Studio are deployed and maintained, not screenshots.
Exploring
- Golden-query evals for text-to-SQL — execution accuracy, not vibes.
- Model Context Protocol (MCP) for typed tool-use across agents.
- Long-context vs. retrieval — when a bigger window still loses to a small, scoped prompt.
- Agent memory beyond flat conversation summaries.
Reading
- The Model Context Protocol specification — typed tool-use as a protocol, not per-agent glue.
- Spider (Yu et al., 2018) — the cross-domain text-to-SQL benchmark I'm measuring DBWhisper against.
Open to
Open to AI software engineering roles — remote or Ahmedabad, India.