Skip to content
Technology

Retrieval-Augmented Fine-Tuning (RAFT)

A training technique that combines retrieval-augmented generation with fine-tuning, teaching models to better leverage retrieved context. Produces models that are both knowledgeable and grounded in source documents.

Why it matters

RAG lets a model pull in outside documents at answer time, but models aren't always good at actually using what's retrieved, and sometimes ignore it. RAFT trains the model specifically to read retrieved passages well and lean on them instead of guessing. For teams building AI over their own knowledge base, this means more answers grounded in the real documents and fewer confident-sounding mistakes.

In practice

A company builds an internal assistant over its policy manuals. With plain RAG, the model sometimes retrieves the right page but still answers from memory. After RAFT, it's been trained on examples where the correct answer comes straight from the retrieved text, and also to ignore irrelevant passages. Now when an employee asks about a leave policy, the reply reliably reflects what the actual manual says.

Related terms

Put Retrieval-Augmented Fine-Tuning (RAFT) into practice

Access 800+ AI models and 70+ tools through Vincony — start free with 100 credits.