Is Fine-Tuning Still Needed? LLMs, RAG, & LoRA
127K views · Jul 21, 2026 · Education
Comments · 165
@遊蕩者 · 2 months ago
distillation-oriented finetuning may be the most popular trend in community, and Chinese AI labs.<br>As long as tasks are simple enough, we can use smart small model to replace large frontier one.
11
@ttt6262 · 2 months ago
I definitely think fine-tuning is still very relevant, especially for smaller models with specialized tasks. Of course, the very large models are significantly better in every respect than these smaller models, which is precisely why fine-tuning is so important for smaller models—to ensure they perform their task at least as well as the very large models, while running locally and offering significantly better data protection than the large cloud models.
27
@jeffreywhewhetu · 2 months ago
With what I just learnt from the video? I think privacy would be one of the major reasons to do fine tuning if you ask me<br>Thank you for updating us on the state of AI via your videos, IBM!!
42
@chrismoore4803 · 2 months ago
100% agree with what we’ve seen in the field. The one Case for fine-tuning that I have also seen do well is capturing tone which is used in a specific industry. Your legal example may also be suggestive of this phenomenon. You mentioned that lawyers liked the output of the fine-tune model. Clearly that doesn’t necessarily make the model quantitatively better or worse, but we have seen that fine-tuning captures. The tone of voice used in a field, and it seems to resonate with experts in a way that prompt guidance struggles to capture.
10
@equan18 · 2 months ago
These videos are fantastic, and the best education series on AI! Keep these coming!
23
@RealMcDudu · 2 weeks ago
The video is probably referring to Harvey’s legal AI system. Newer general-purpose frontier models eventually outperformed Harvey’s original specialized model - but that’s not the same as showing that general models outperform a newly fine-tuned model based on the same generation of technology. Specialization may still provide an advantage; the problem is that frontier models are improving so quickly that this advantage can close within a year or even less.
2
@SergeZIEHI · 1 month ago
So mindful & precious time grasping this unvaluable content, delighting every second... One of the best if not the best Nerd content out there.
1
@felix5310 · 2 months ago
I found finetuning to also be very useful if you want to force a certain writing style. Giving examples also works, but if you want it to become really consistent, nothing beats a good finetuning run.
2
@Tenebrisuk · 2 months ago
Fine tuning still has a place for other tasks such as classification and sentiment analysis, where you are likely going to have a more consistent and far more efficient pipeline, not to mention the impacts in terms of compute (which relate to both financial and environmental cost). Not every task is best suited with a generative decoder model. Even the retrieval augmented generation, the retrieval part can occasionally benefit from finetuning.
45
@gorairakesh · 2 months ago (edited)
Also, sometimes a fresh mind helps in problem solving. This made me remember an incident where an IT architect and a team of specialists couldn't resolve an issue and finally they had to involve an external consultant who didn't have any prior knowledge of that specific project. But he was able to resolve the issue on account of his experience in other similar projects. So, in this context I think the base Frontier models bring freshness of mind, who can think deep and independently.
1
@360captureit · 2 months ago
As always, Martin thought-provoking - thank you
@NikhilGoyal-k9m · 3 weeks ago
exceptional! thank you!
Up next

Is RAG Still Needed? Choosing the Best Approach for LLMs
IBM Technology · 1M views

2026 Cost of a Data Breach Report: AI Is Changing Cybersecurity
IBM Technology · 38K views

5 AI Myths & The Truth Behind Them: ML, Context, Agents & More
IBM Technology · 32K views

AI Infrastructure Explained (GPUs, vLLM, and LLM-D)
KodeKloud · 326K views

The E-Commerce Retention Playbook: How to Increase LTV & Repeat Purchases
Mage Loyalty · 50 views

RAG vs. CAG: Solving Knowledge Gaps in AI Models
IBM Technology · 640K views

Deep Dive into LLMs like ChatGPT
Andrej Karpathy · 9.8M views

Stanford CS229 I Machine Learning I Building Large Language Models (LLMs)
Stanford Online · 2.8M views

OWASP's Top 10 Ways to Attack LLMs: AI Vulnerabilities Exposed
IBM Technology · 312K views

Llama.cpp vs vLLM: Which Local LLM Engine Actually Scales?
IBM Technology · 95K views

AI Agents For Beginners – OpenClaw Case Study
freeCodeCamp.org · 161K views

AI Model vs Agentic Harness: What Actually Drives AI
IBM Technology · 144K views

Prompting Is Dead in 6 Months. Andrew Ng, Stanford
philia · 225K views

You Can Learn AI Agent Harness & Loop Engineering In 19 Min | LLM Ops, Eval, Tracing, RAG
Sean‘s AI Stories and Waku Agent · 268K views

Agent Harness explained in 8min..
Caleb Writes Code · 610K views

RAG vs Fine-Tuning vs Prompt Engineering: Optimizing AI Models
IBM Technology · 730K views

RFT, DPO, SFT: Fine-tuning with OpenAI — Ilan Bigio, OpenAI
AI Engineer · 20K views

The 7 Skills You Need to Build AI Agents
IBM Technology · 566K views

What AI Agent Skills Are and How They Work
IBM Technology · 473K views

AI-103 Develop AI Apps and Agents on Azure Study Cram
John Savill's Technical Training · 43K views

When to Build Your Own Agent Harness | Harrison Chase, LangChain
Sequoia Capital · 95K views

Ilya Sutskever – We're moving from the age of scaling to the age of research
Dwarkesh Patel · 1.5M views

5 AI Agent Terms You Need to Know
IBM Technology · 113K views

Don't learn AI Agents without Learning these Fundamentals
KodeKloud · 1.2M views

Scott Jenson: Are we really going to use the same Desktop UX forever?
The KDE Community · 419K views

Transformers, the tech behind LLMs | Deep Learning Chapter 5
3Blue1Brown · 11M views