Context Windows, Chunking and Compression: Feed the Model Right
by Priya Raman
Catalog
Programming, AI and prompt-engineering courses written by working engineers. Filter by language, level, price or rating — buy once, get an instant email access link, keep it for life.
Jump to a technology
Showing 1–7 of 7 courses
Fit the right material into a context window and cut token spend without losing accuracy.
Understand embeddings well enough to build search that finds meaning, not just keywords.
Chunking, hybrid search and reranking, the parts that decide whether retrieval really works.
Pick, index and operate a vector store without turning your search into a science project.
Build a dataset, run a LoRA fine-tune, and prove it beats prompting for your narrow task.
Queues, caching, streaming, fallbacks and budgets: the architecture behind a reliable AI feature.
Twenty-four hours taking you from a first API call to a deployed, evaluated AI product.