Janu Verma

Janu Verma

I study how machines learn, remember, and reason. I strive to understand things deeply, and to share what I find. Over the last few years my focus has moved towards building learning systems that reason, adapt, and improve through feedback. Currently I am a principal applied scientist working on AI personalization at Microsoft, in London.

I trained in mathematics and physics, and a lot of my time still goes to them: reading, working through problems, and following science more generally. I have spent twelve years turning research into systems people use: in genomics, healthcare, payments, social media, and now agentic AICornell, IBM Research, Mastercard, Hike, Cult.fit, Microsoft. Cambridge and Kansas State before that. The longer version is below.. I write Incomplete Distillation, a research journal, and build small, complete systems to check what I think I know.

Now

Working on
Personalization for Copilot by day. In public, the physics of coffee brewing, one method at a time: espresso and pour-over are written up, and the wet bed is still the open problem.
Reading
NW, Zadie Smith·The Pearl, John Steinbeck·Orbital, Samantha Harvey
Top of mind
The cosmological constant and the heat death of the universe·Type theory and the Curry–Howard correspondence·Temporal event sequences

as of September 2026

Projects all projects →

Also: Instagram Strategist, a content-strategy agent on Gemma 4. The projects group into three tracks: Personal Intelligence, RecSys, Explorable Explanations.

Writing januverma.substack.com →

Incomplete Distillation is a research journal: long pieces, usually with code and a small experiment behind them, on whatever I am trying to understand that monthThe name is a confession. Each essay is a partial distillation of a subject that deserves more; a series continues until the residue is small enough.. Thirty-four essays since June 2025, mapped below by thread.

Jul 2025 Oct Jan 2026 Apr Jul Recommendation & personalization 9 2025-06-30 · LLM-based Cross-encoder for Recommendation Systems 2025-08-04 · Semantic IDs: A Technical Deep Dive 2025-08-29 · Semantic IDs: A Practical Study 2025-09-15 · Contextual Bandits in Recommender Systems 2025-10-03 · Joint Retrieval and Recommendation Modeling 2025-10-05 · RecSys 2025 Recap 2026-01-17 · RecSys After LLMs: Four Paradigms 2026-05-26 · Marginalia 2026-06-04 · Personal Intelligence Agents & reinforcement learning 7 2025-10-12 · Fine-tuning LLMs using RL 2025-10-19 · Fine-tuning LLMs using RL II: From RLOO to GRPO 2025-11-03 · Multi-Turn Tool Use with RL 2025-12-13 · LLM-based Agents 2026-01-09 · Building a Data Analysis Agent 2026-01-25 · From Custom Parsing to MCP 2026-04-14 · I Gave Gemma 4 My Instagram Reels Models & architectures 9 2025-07-17 · Graph Transformers 2026-02-07 · Graph Foundation Models 2026-02-28 · Diffusion Models I 2026-03-06 · Diffusion Models II 2026-03-14 · Diffusion Models III 2026-03-21 · Research Briefings: Video-JEPA 2.1 2026-03-28 · Research Briefings: Mamba-3 2026-04-07 · Research Briefings: TurboQuant 2026-04-11 · Research Briefings: Gemma 4 AI × science 5 2025-07-08 · The Protein Folding Problem 2025-07-23 · From Silver to Gold: AI at the IMO 2025-09-07 · Protein Design: A Toy Case Study 2026-08-07 · Physics of Coffee Brewing I 2026-09-11 · Physics of Coffee Brewing II: Pour-over Methods Essays 4 2026-01-02 · On New Year's Resolutions 2026-02-01 · Game-Theoretic Thinking for AI 2026-02-09 · Claude Built My Website 2026-02-15 · What a Time to Be Alive
Every essay since June 2025, by thread; thirty-four so far. Hover for the title, click to read. The red mark is the newest.
Recommendation & personalization9 essays
Agents & reinforcement learning7 essays
Models & architectures9 essays
AI × science5 essays
Essays4 pieces

Research google scholar →

Research has been my way into each field: crop genomics at Cornell, machine learning for healthcare at IBM Research, graph learning and semi-supervised methods for fraud at Mastercard, recommendation and personalization at Hike and Microsoft. Much of it shipped rather than published. Selected papersAlso patents: AI methods for predicting account-level risk of cardholders (Mastercard, filed 2022 and 2025), and mechanism-of-action derivation for adverse drug reaction prediction (IBM, 2019).:

  1. Generative Recommenders for Zero-Query Prompt Recommendation. MLADS 2025 with V. Kolesnyk, R. Ronen
  2. A Closer Look at Consistency Regularization for Semi-Supervised Learning. CODS-COMAD 2024 S. Ghosh, S. Kumar, A. Kumar, J. Verma
  3. Guided Self-Training Based Semi-Supervised Learning for Fraud Detection. ICAIF 2022 A. Kumar, S. Ghosh, J. Verma
  4. Label-aware Sampling using Contrastive Learning for GNN-based Fraud Detection. KDD 2022, Machine Learning in Finance J. Verma, G. Arora, A. Patankar, A. Chaudhry
  5. Heterogeneous Edge Embedding for Friend Recommendation. ECIR 2019 J. Verma, S. Gupta, D. Mukherjee, T. Chakraborty
  6. Cassava Haplotype Map Highlights Fixation of Deleterious Mutations during Clonal Propagation. Nature Genetics 2017 P. Ramu et al.

Speaking and advising invite me →

I speak about personalization and recommendation, agentic AI, and what separates robust AI systems from demos. Open to guest lectures and teachingAlso spoken at IIIT Delhi, University of Delhi, and the Institute of Mathematical Sciences, Chennai..

  • 2026·05Preparing for an AI-First WorldKeynote · International Conference on Manoeuvring Business, Society and Culture, Jammu
  • Challenges in Retrieval for Enterprise AI AgentsGuest lecture · Applied AI, Rutgers University
  • LLMs in Recommender SystemsPyData London
  • Graph Embedding Methods for RecommendationData Hacks Summit, Analytics Vidhya
  • Transfer Learning in NLPPyData Delhi
  • Beyond QWERTY: Solving the Input Problem of IndiaFuture of Work, YourStory

I also advise founders on getting from an AI concept to a testable, dependable product: technical advisor to sia.vision, London, on agentic systems for storytelling and visual media. Open to advising UK founders working on agentic AI or personalization.

About

I started where the abstractions are purest: Cambridge, studying the mathematics of string theory and geometryMMath, Part III of the Mathematical Tripos, with a dissertation applying gauge/gravity duality to fluid dynamics. Then an MS in Mathematics at Kansas State, on geometric invariants arising from supersymmetric quantum field theories, and three years of a PhD before leaving to build things.. That search took me through computational biology at Cornell, drug discovery and clinical NLP at IBM Research, graph-based fraud and risk models at Mastercard, where I led a team of scientists and engineers, and personalised recommendation at Hike, Cult.fit, and now Microsoft.

The maths never left. It shaped how I think about representation, structure, and abstraction, and it shows up as a need to actually understand something before I trust it. The domain changes. The way of thinking does not. Writing is how I hold myself to that: putting an idea down in full, with the code beside it, is the fastest way I know to find out whether I understood it.

Off duty: pour-over videos, recipes, pop culture, and thinking too much about what to wear.