How does a text watermark work?
A first-principles investigation of how generation-time text watermarks work, from a weighted coin to Gemma, and where the evidence stops.
I am a Senior Data Scientist at 6sense building production AI applications, foundational model systems, custom embeddings, and agent evaluation frameworks. On this site, I write in-depth technical essays documenting what I build, break, and learn across large language models, retrieval-augmented generation (RAG), statistical text watermarking, and autonomous developer agents.
A first-principles investigation of how generation-time text watermarks work, from a weighted coin to Gemma, and where the evidence stops.
Testing if applying Matryoshka Representation Learning (MRL) to tabular entity data could bridge the cost gap, compressing embeddings enough to remain operational while still beating BM25 on corrupted queries.

A production diary of tiered episodic memory in AI agents. Three markdown files, an SQLite database, and a lobster that somehow remembers what you said three days ago.

Four tools, a settings file, and full control - how I built my own AI coding agent with Pi instead of paying $200/month for a CLI that keeps changing.

Practical way of using colbert with ragatouille on modal labs

Practical strategies for implementing and optimizing Retrieval Augmented Generation (RAG) in LLM systems.

Hackathon Experience at Mistral Hackathon in San Francisco



