field notes ยท zero fluff

writings

what i learn, as i learn it โ€” evals, agents, rag, reasoning. written for people who ship.

11 notes

01 three flavors of simulating an agent pipeline
02 evolving an eval harness from one bash script to a merge gate
03 chunk twice, retrieve once: double-pass semantic chunking
04 building a context system that agents can actually use
05 teaching a small llm to filter its own context
06 grpo + lora: making a 3b model reason
07 colqwenrag: rag without ocr, for legal papers
08 fine-tuning embeddings until legal-rag stops missing
09 rolling your own reasoner: grpo + sft from scratch
10 how to make a model think: the deepseek recipe
11 a llm that drafts indian legal contracts