Chunking is the least glamorous and most decisive knob in RAG, agreed. What finally worked for me was chunking on document structure instead of a fixed token count, then scoring retrieval against a labeled set so chunk size became a measured decision and not a guess. Are you tuning chunk size by eye, or do you have a retrieval metric you watch when you change it?