Articles for people who
ship the stack
Original rewrites on React, Node.js, TypeScript and AI — practical notes from the same engineering practice behind our operator software. Article bodies are in English.
Tagged: nlp
Prompt Engineering: Steer LLMs Without Fine-Tuning
Use roles, few-shot examples, chain-of-thought, and format constraints to guide models — and know when prompting alone hits a ceiling versus fine-tuning.
748 wordsRead articleChoosing and Tuning Embedding Models for Production RAG Systems
Learn how embedding models turn text into searchable vectors, why domain vocabulary breaks semantic search, and how to select, compress, and fine-tune models for production RAG.
6591 wordsRead articleRetrieval-Augmented Generation Explained: Fixing LLM Knowledge Gaps
Learn why LLMs hallucinate and go stale, then see step-by-step how RAG retrieves, chunks, embeds, and augments prompts to fix it.
1093 wordsRead articleReducing Hallucinations in a Medical RAG Chatbot Pipeline
Learn how hybrid search, reranking, and a strict refuse-to-fabricate policy combine to build a more trustworthy medical research RAG chatbot.
1960 wordsRead articleSecond Brain: Turning Meeting Transcripts into a Queryable Knowledge Graph
Explains how an agentic system extracts entities from meeting transcripts and stores them in Cosmos DB to enable natural-language recall and knowledge graph exploration.
1390 wordsRead article
About these articles
Request a 24h estimate
Need the same stack in a production operator layer? Send the brief — estimate within 24 hours.