From e71cf2b9319e9471d26d6ca5fed1b6d0aee6b623 Mon Sep 17 00:00:00 2001 From: =?UTF-8?q?Jukka=20Sepp=C3=A4nen?= <40791699+kijai@users.noreply.github.com> Date: Thu, 19 Dec 2024 14:02:02 +0200 Subject: [PATCH] Update readme.md --- readme.md | 2 +- 1 file changed, 1 insertion(+), 1 deletion(-) diff --git a/readme.md b/readme.md index 7e9f7ca..fc7f3b8 100644 --- a/readme.md +++ b/readme.md @@ -9,7 +9,7 @@ So - very much like IPAdapter - but VLM will do the heavy lifting for you! Now this is a tuning free approach but with further task specific tuning we can expand the use scenarios. -# Guide to Using `xtuner/llava-llama-3-8b-v1_1-transformers` for Image-Text Tasks +## Guide to Using `xtuner/llava-llama-3-8b-v1_1-transformers` for Image-Text Tasks ## Step 1: Model Selection Use the original `xtuner/llava-llama-3-8b-v1_1-transformers` model which includes the vision tower. You have two options: