How to work with me
How to reach Soumendra Kumar Sahoo and write a first message that gets a useful reply: what to include, what to expect and what he will politely decline. Two emails arrive on the same morning. The first, subject “Quick question”, hopes I’m doing well and asks whether I’d have time to discuss observability. The second asks whether head-based sampling drops LLM spans at 10%. I answer the second from my phone before the coffee is done. The first gets a reply asking what the question actually is, and we lose two days to that round trip. Nothing about my calendar or my willingness to help changes between those two replies. Only the message does. Put your real question, with enough context to answer it, in the first message. Everything below is detail on that. Messages come from readers, collaborators, recruiters and people stuck on a problem. I like getting them. Email. My address is in the footer of every page on this site, or via the contact section. I read every note and reply to the ones I can help with, usually within a few days. If a week passes with no reply, I’m overloaded, not offended. A polite nudge is welcome and it works. Make your first message complete enough that I can answer it without asking you anything. Hi Soumendra, hope you're doing well! I had a question about LLM observability. Would you be open to connecting? Hi Soumendra! Quick one: does head-based sampling drop LLM spans too, or are they exempt? We sample at 10% and I can't tell if our missing generations are a bug. No rush. Fine, I use these tools too. But edit it down to your words and your actual situation before you send it. If you paste raw model output, you’re sending me something I could have generated myself in four seconds, and a long unedited AI message signals that you didn’t invest the effort you’re asking of me. I hope this email finds you well. I am reaching out because… Hi Soumendra, I drafted this with Claude, then cut it to what matters: I'm building an agent that routes support tickets, my evals disagree with production outcomes and I've tried recalibrating thresholds. Is eval-driven routing even the right direction here? I work async-first, in IST (UTC+5:30), and I think best in writing, so a well-written email usually gets you a better answer than a call would. If something genuinely needs a call, propose it in the email with the agenda and the outcome you want, and I’ll share times. Hi Soumendra, I'm hitting a wall instrumenting our RAG pipeline and would love to walk you through it live. Are you free Tuesday or Wednesday? Hi Soumendra, I'm adding tracing to a RAG pipeline (LangChain + GPT-4). Spans show up for retrieval but not for the generation step; I've tried X and Y from the docs. Minimal repro: (link). Am I missing an instrumentation hook? No rush, whenever you've got ten minutes. A few things about how I communicate that can read wrong: To save us both a round trip: None of this is original. I collected it over the years from people who wrote it down first: I don’t always live up to this page. If your experience of working with me doesn’t match it, tell me. That is exactly the kind of message I want.How to reach me
Put everything in the first message
If AI helped you write it
Calls and meetings
Don’t misunderstand me
What I’ll politely decline
Where this comes from