Skip to content

Fine-Tuning

Last updated:

What fine-tuning is

Large language models start with general pre-training on huge amounts of text. Fine-tuning is a second, much smaller training run on examples you choose, usually pairs of inputs and the outputs you want. The result is a new version of the model whose internal parameters, called weights, have shifted toward those examples.

Fine-tuning is good at teaching patterns: a consistent tone, a fixed output format, a classification scheme or domain-specific phrasing. It is less suited to teaching facts that change, such as prices, stock or opening hours, because every change would need another training run.

How fine-tuning works

A typical fine-tuning project follows four steps:

  • Collect examples: a set of high-quality input and output pairs that show the behavior you want. Quality and consistency matter more than volume.
  • Train: the model provider or your own infrastructure adjusts the model’s weights on those examples, often with methods such as LoRA that update only a small part of the model.
  • Evaluate: compare the tuned model with the original on test questions it has not seen, to check that it improved without losing general ability.
  • Deploy and maintain: the tuned model replaces the base model in your application and has to be retrained when the base model is retired or the desired behavior changes.

Example

A law firm wants its internal assistant to summarize client calls in a strict five-part format. Instructions alone get the format right most of the time, but not always. The firm fine-tunes a model on a few hundred past summaries written by its staff, and the tuned model follows the format more consistently. For anything factual, such as the current fee schedule, the firm still relies on retrieval, because those details change during the year.

Why it matters for business chatbots

For most website chatbots, the main job is answering accurately from the company’s current content. Retrieval-augmented generation (RAG) usually handles that better than fine-tuning: updates take effect once content is re-indexed, answers can be traced to sources, and there is no training data set to build. Fine-tuning is worth considering when a business needs a specific behavior that clear instructions and examples cannot achieve, and has the data and budget to maintain a custom model.

Fine-tuningRAG
What changesThe model’s weightsThe text given to the model
Best forStyle, format, narrow tasksFacts, policies, product information
Updating knowledgeTrain the model againEdit and re-index the content
TraceabilityHard to tell where an answer came fromAnswers can point to source passages

Fine-tuning and intoCHAT

intoCHAT does not fine-tune models. When you “train” an intoCHAT agent, it processes your website pages, files, text snippets and Q&A pairs into searchable chunks with embeddings. The underlying model, one OpenAI model chosen and managed by intoCHAT, stays the same. That is why changes take effect after you retrain a source, without waiting for a training run. To shape tone and behavior, you write the agent’s instructions and add Q&A pairs for answers that must be exact.

Frequently asked questions

Is fine-tuning the same as training a chatbot on my website?

Usually not. Most website chatbot platforms, intoCHAT included, “train” on your content by indexing it for retrieval, not by changing the model’s weights. Fine-tuning is a separate, more involved process that creates a modified version of the model.

Should I use fine-tuning or RAG for a customer support chatbot?

For questions about products, policies and prices, RAG is usually the better choice, because the knowledge stays easy to update and check. Fine-tuning can help with a consistent style or a narrow task, and some teams combine both. Start with retrieval and clear instructions, and consider fine-tuning only if a clear gap remains.

Does fine-tuning stop hallucinations?

No. A fine-tuned model can still give confident wrong answers, especially about facts that were not in its training examples or have changed since. Grounding answers in retrieved sources and letting the model say when it does not know are more direct ways to reduce that risk.

Does intoCHAT fine-tune a model on my data?

No. intoCHAT indexes your content for search and passes relevant excerpts to the model with each question; it does not train or fine-tune a model on your data. You control behavior through instructions, Q&A pairs and the sources you add.

See it answer from your own website

Paste your website address and chat with an agent built from your pages. It takes about a minute.

Create your agent free

Free plan, no credit card needed.