Search
Search titles only
By:
Search titles only
By:
Log in
Register
Search
Search titles only
By:
Search titles only
By:
Menu
Install the app
Install
Forums
New posts
All threads
Latest threads
New posts
Trending threads
Trending
Search forums
What's new
New posts
New ads
New profile posts
Latest activity
Free Ads
Latest reviews
Search ads
Members
Current visitors
New profile posts
Search profile posts
Contact us
Latest ads
Premium Land with House for Sale
anil1961
Updated:
Friday at 10:15 AM
AWS Certified Solutions Architect-Associate + AWS Certified Cloud Practitioner
Sanjeewani95
Updated:
Wednesday at 8:16 PM
🚀 එක පැකේජ් එකයි - මාසෙටම Unlimited Internet! 🌐
sayuru bandara
Updated:
Tuesday at 10:57 AM
🎬 CapCut Pro 1 Month Access! LKR 600
sayuru bandara
Updated:
Tuesday at 10:55 AM
🚀 Google One AI PRO Plan (Gemini Pro Activation) – 18 Months Access! LKR 2200
sayuru bandara
Updated:
Tuesday at 10:53 AM
Electronics
Vehicles
Property
Search
Reply to thread
Forums
General
ElaKiri Talk!
බ්ලූ ටික් අයියාලගේ ඒඅයි වික්රම. RUN LLAMA Model Locally
Get the App
JavaScript is disabled. For a better experience, please enable JavaScript in your browser before proceeding.
You are using an out of date browser. It may not display this or other websites correctly.
You should upgrade or use an
alternative browser
.
Message
<blockquote data-quote="MihiCherub" data-source="post: 30871139" data-attributes="member: 238676"><p>FB එකේ බ්ලූ ටික් අයියාලා මේ ටිකේ සෑහෙන්න ටෝක් කරනව උන් AI හැදුවා අරකයි මේකයි ගගා. අපේ සින්හලුත් වැලලෙනව "නියමයි අයියාහ්" කියල. උන් ඉතින් එතන ඉදන් උපරිම ආස්වාදෙන් පට්ට කෙබර ලෝකයක් කෙලිනව තව තව ලොකු කතාත් එමට කියනව. සමහරු නම් උන් කියන ඒව උන් දන්නෙත් නෑ. සමහරු නම් chatgpt එකටත් ඌ හදපු එකෙන් කෙලවලා දාන්න ටෝක් ඉශූ කරන්නෙ. උන්ටම වන්දනාමාන කරන ගෝත්රිකයො එක්ක කතා කරල වැඩකුත් නෑ. අපි හෙන ලොකු සුපිරි AI එකක් හැදුවා කියන ගන්කබරයො ඔක්කොම කරන්නෙ chatgpt, claude, gemini වගේ මොඩල් එකක API keys අරන් integrate කරන එක. ඊට පස්සෙ උගේ platform එකක back end එකෙන් api access කරන එක. රෑ එලිවෙනකම් දිවා රෑ නොබලා මාස ගානක ප්රතිපල කියල හෑල්ලක් එක්ක තමා ඕක දාන්නෙ. ඕකත් AI එකටම කියල හදාගත්ත නම් ඉවරයි.</p><p></p><p>ඔතනින් සුලු පිරිසක් ඉන්නව agents build කරගෙන තව ටිකක් fine tune කරපු output එකක් දෙන. හැබැයි මම නම් මේ වෙනකම් හරියට perform කරන finetune කරපු model එකක් දැකල නෑ. තියෙනව නම් model card එකක් දාන්න බලන්න. ඒ වගේම බ්ලූ ටික් හෑලි අයියලා කිසිම එකෙක් අඩුම ගාන instance එකක් rent කරල මොඩල් එකක් රන් කරලවත් උන්ගෙ platform එකක් හදන් නෑ ටෝක් ඉශූ කරාට.</p><p></p><p>මේකෙ තියෙන ලොකුම ප්රශ්නෙ මෙහෙම බොරුමවාපෑම් කරල කෝස් කරන්න මිනිස්සු රවට්ටන එක. prompt engineering කියලත් කෝස් කරනව. මාත් free session එකකට join වෙලා බැලුව. හාල්පාරුව නිකන් මේ වචන ගගහ output ගන්න කියනව මිසක් ඌ මොඩල් කාඩ් එකේ technical paper එකවත් කියවල නෑ. ඉතින් මේ AI මැජික් කරන් සින්හලු අනාත කරන නිසා මේක ලිව්වෙ.</p><h2><span style="font-size: 22px">RUN LLAMA Model Locally</span></h2><p>ඔන්න කැමති ඕනෙ එකෙක් ඉන්නව තමන්ගෙම PC එකක locally run කරගනිල්ල.</p><p>දැනට released වෙලා තියෙන models වලින් censored / uncensored කියල කොටස් දෙකක් තියෙනව. censored කියන්නෙ chatgpt, gemini, claude, sonnet වගේ අපිට ඕනෙ හැමදෙයක්ම කරගන්න බෑ limitation තියෙනව. uncensored model වල එහෙම දෙයක් නෑ. explicit content වල ඉදන් ඕනෙම දෙයක් අහන්න පුලුවන් කිසිම limitation එකක් නෑ. safety guardrails, content filters, or refusal mechanisms මුකූත් නෑ.</p><p></p><p>1.1 මුලින්ම python install කරගන්න.</p><p><a href="https://www.python.org/downloads/release/python-31011/" target="_blank">https://www.python.org/downloads/release/python-31011/</a></p><p><a href="https://www.python.org/ftp/python/3.10.11/python-3.10.11-amd64.exe" target="_blank">https://www.python.org/ftp/python/3.10.11/python-3.10.11-amd64.exe</a></p><p></p><p>1.2 Install git</p><p><a href="https://git-scm.com/downloads" target="_blank">https://git-scm.com/downloads</a></p><p><a href="https://github.com/git-for-windows/git/releases/download/v2.50.1.windows.1/Git-2.50.1-64-bit.exe" target="_blank">https://github.com/git-for-windows/git/releases/download/v2.50.1.windows.1/Git-2.50.1-64-bit.exe</a></p><p></p><p>1.3 CMD open කරගන්න</p><p></p><p>2. Update Your System:</p><p>sudo gives you administrative privileges, and -y automatically confirms prompts.</p><p></p><p>[CODE=bash]sudo apt update</p><p>sudo apt upgrade -y[/CODE]</p><p></p><p>3. Verify NVIDIA Driver Installation:</p><p>table එකක් එන්න ඕනෙ එකේ gpu nvidia version එකයි cuda වර්ශන් එකයි. drivers නැත්නම් drivers ටික install කරගන්න</p><p>[CODE=bash]nvidia-smi[/CODE]</p><p></p><h3>Windows</h3><p>4.1 Download & Install Ollama</p><p>මේකෙන් තමයි models run කරගන්නෙ.</p><p><a href="https://ollama.com/download/windows" target="_blank">https://ollama.com/download/windows</a></p><p></p><p>4.2 Check Ollama version. ex ollama version is 0.9.6</p><p>[CODE=bash]ollama --version[/CODE]</p><p></p><h3>Linux</h3><p>4.1 Download & Install Ollama</p><p>[CODE=bash]curl -fsSL https://ollama.com/install.sh | sh[/CODE]</p><p></p><p>4.2 Check Ollama version. ex ollama version is 0.9.6</p><p>[CODE=bash]ollama --version[/CODE]</p><p></p><p>5. Download model</p><p>ඕන model එකක් මෙතනින් search කරගන්න.</p><p><a href="https://ollama.com/search" target="_blank">https://ollama.com/search</a></p><p></p><p>small model එකකින් පටන් ගමු.</p><p></p><p>5.1 gemma3:1b </p><p><a href="https://ollama.com/library/gemma3:1b" target="_blank">https://ollama.com/library/gemma3:1b</a></p><p>මේක lightweight censored model එකක්. single gpu එකක run කරගන්න පුලුවන්. 1B parameters තියෙනව. text input කරන්න පුලුවන්. 32k context window. 140 languages</p><p></p><p>[CODE=bash]ollama pull gemma3:1b [/CODE]</p><p></p><p>5.2 gemma3:4b (recommended model) </p><p><a href="https://ollama.com/library/gemma3:4b" target="_blank">https://ollama.com/library/gemma3:4b</a></p><p>මේක lightweight censored multimodal එකක්. single gpu එකක run කරගන්න පුලුවන්. 4B parameters තියෙනව. image/text දෙකම input කරන්න පුලුවන්. 128k context window. 140 languages</p><p></p><p>[CODE=bash]ollama pull gemma3:4b [/CODE]</p><p></p><p>6. List Downloaded Models</p><p>[CODE=bash]ollama list[/CODE]</p><p></p><p>7. Run Your First Model</p><p>[CODE=bash]ollama run gemma3:1b[/CODE]</p><p></p><p>8. To exit the interactive session, type /bye or press Ctrl + D</p><p></p><p>9. Install openweb ui</p><p>9.1 Create directory and navigate</p><p>[CODE=bash]</p><p>mkdir ~/open-webui-venv</p><p>cd ~/open-webui-venv</p><p>[/CODE]</p><p></p><p>9.2 Create a Python Virtual Environment</p><p>[CODE=bash]python3 -m venv venv[/CODE]</p><p></p><p>9.3 Activate the Virtual Environment</p><p>[CODE=bash]source venv/bin/activate[/CODE]</p><p>You'll notice your terminal prompt changes (e.g., (venv) username@hostname:~/open-webui-venv$) to indicate that the virtual environment is active.</p><p></p><p>9.4 Install Open WebUI</p><p>[CODE=bash]pip install open-webui[/CODE]</p><p></p><p>10.1 Start the Open WebUI Server</p><p>[CODE=bash]open-webui serve[/CODE]</p><p>This will usually start the server on <a href="http://localhost:8080" target="_blank">http://localhost:8080</a>. You can then access it from your Windows browser</p><p></p><p><a href="http://localhost:8080" target="_blank">http://localhost:8080</a></p><p></p><p>10.2 Deactivate the Virtual Environment (when you're done)</p><p>[CODE]deactivate[/CODE]</p><p></p><h4>Uncensored Qwen3 - JOSIEFIED 8B</h4><p><a href="https://ollama.com/goekdenizguelmez/JOSIEFIED-Qwen3" target="_blank">https://ollama.com/goekdenizguelmez/JOSIEFIED-Qwen3</a></p><p>[CODE=bash]ollama run goekdenizguelmez/JOSIEFIED-Qwen3:8b[/CODE]</p><p>This model has reduced safety filtering and may generate sensitive or controversial outputs. Use responsibly and at your own risk.</p><p></p><h4>Using Multimodal Models (LLaVA with Images)</h4><p>[CODE=bash]ollama run llava "What do you see in this image? /home/yourusername/Pictures/my_image.jpg"[/CODE]</p><p>(Replace /home/yourusername/Pictures/my_image.jpg with the actual path to your image.)</p><p></p><p>මේ විදියට රන් කරගන්න පුලුවන් gpt, gemini වගේම models තියෙනව</p><p>qwen3 කියන්නෙ එහෙම හොද කොලිටි මොඩල්ස්. 142GB 40K context, grok, deepseek r1, gemini 2.5, openai-o1 වලට වඩා benchmark score එක හොදයි.</p><p><a href="https://ollama.com/library/qwen3:235b" target="_blank">https://ollama.com/library/qwen3:235b</a></p><p></p><p>code කරන්න qwen coder එහෙම සුපිරි. OpenAI’s GPT-4oට වඩා හොදයි.</p><p><a href="https://ollama.com/library/qwen2.5-coder" target="_blank">https://ollama.com/library/qwen2.5-coder</a></p><p></p><p>ප්රශ්න මොනාහරි තියෙනව නම් දාපල්ල. gpu එකක් නැත්නම් gpu instance එකක රන් කරගන්න help එකක් ඕනෙ නම් කියපල්ල.</p></blockquote><p></p>
[QUOTE="MihiCherub, post: 30871139, member: 238676"] FB එකේ බ්ලූ ටික් අයියාලා මේ ටිකේ සෑහෙන්න ටෝක් කරනව උන් AI හැදුවා අරකයි මේකයි ගගා. අපේ සින්හලුත් වැලලෙනව "නියමයි අයියාහ්" කියල. උන් ඉතින් එතන ඉදන් උපරිම ආස්වාදෙන් පට්ට කෙබර ලෝකයක් කෙලිනව තව තව ලොකු කතාත් එමට කියනව. සමහරු නම් උන් කියන ඒව උන් දන්නෙත් නෑ. සමහරු නම් chatgpt එකටත් ඌ හදපු එකෙන් කෙලවලා දාන්න ටෝක් ඉශූ කරන්නෙ. උන්ටම වන්දනාමාන කරන ගෝත්රිකයො එක්ක කතා කරල වැඩකුත් නෑ. අපි හෙන ලොකු සුපිරි AI එකක් හැදුවා කියන ගන්කබරයො ඔක්කොම කරන්නෙ chatgpt, claude, gemini වගේ මොඩල් එකක API keys අරන් integrate කරන එක. ඊට පස්සෙ උගේ platform එකක back end එකෙන් api access කරන එක. රෑ එලිවෙනකම් දිවා රෑ නොබලා මාස ගානක ප්රතිපල කියල හෑල්ලක් එක්ක තමා ඕක දාන්නෙ. ඕකත් AI එකටම කියල හදාගත්ත නම් ඉවරයි. ඔතනින් සුලු පිරිසක් ඉන්නව agents build කරගෙන තව ටිකක් fine tune කරපු output එකක් දෙන. හැබැයි මම නම් මේ වෙනකම් හරියට perform කරන finetune කරපු model එකක් දැකල නෑ. තියෙනව නම් model card එකක් දාන්න බලන්න. ඒ වගේම බ්ලූ ටික් හෑලි අයියලා කිසිම එකෙක් අඩුම ගාන instance එකක් rent කරල මොඩල් එකක් රන් කරලවත් උන්ගෙ platform එකක් හදන් නෑ ටෝක් ඉශූ කරාට. මේකෙ තියෙන ලොකුම ප්රශ්නෙ මෙහෙම බොරුමවාපෑම් කරල කෝස් කරන්න මිනිස්සු රවට්ටන එක. prompt engineering කියලත් කෝස් කරනව. මාත් free session එකකට join වෙලා බැලුව. හාල්පාරුව නිකන් මේ වචන ගගහ output ගන්න කියනව මිසක් ඌ මොඩල් කාඩ් එකේ technical paper එකවත් කියවල නෑ. ඉතින් මේ AI මැජික් කරන් සින්හලු අනාත කරන නිසා මේක ලිව්වෙ. [HEADING=1][SIZE=6]RUN LLAMA Model Locally[/SIZE][/HEADING] ඔන්න කැමති ඕනෙ එකෙක් ඉන්නව තමන්ගෙම PC එකක locally run කරගනිල්ල. දැනට released වෙලා තියෙන models වලින් censored / uncensored කියල කොටස් දෙකක් තියෙනව. censored කියන්නෙ chatgpt, gemini, claude, sonnet වගේ අපිට ඕනෙ හැමදෙයක්ම කරගන්න බෑ limitation තියෙනව. uncensored model වල එහෙම දෙයක් නෑ. explicit content වල ඉදන් ඕනෙම දෙයක් අහන්න පුලුවන් කිසිම limitation එකක් නෑ. safety guardrails, content filters, or refusal mechanisms මුකූත් නෑ. 1.1 මුලින්ම python install කරගන්න. [URL]https://www.python.org/downloads/release/python-31011/[/URL] [URL]https://www.python.org/ftp/python/3.10.11/python-3.10.11-amd64.exe[/URL] 1.2 Install git [URL]https://git-scm.com/downloads[/URL] [URL]https://github.com/git-for-windows/git/releases/download/v2.50.1.windows.1/Git-2.50.1-64-bit.exe[/URL] 1.3 CMD open කරගන්න 2. Update Your System: sudo gives you administrative privileges, and -y automatically confirms prompts. [CODE=bash]sudo apt update sudo apt upgrade -y[/CODE] 3. Verify NVIDIA Driver Installation: table එකක් එන්න ඕනෙ එකේ gpu nvidia version එකයි cuda වර්ශන් එකයි. drivers නැත්නම් drivers ටික install කරගන්න [CODE=bash]nvidia-smi[/CODE] [HEADING=2]Windows[/HEADING] 4.1 Download & Install Ollama මේකෙන් තමයි models run කරගන්නෙ. [URL]https://ollama.com/download/windows[/URL] 4.2 Check Ollama version. ex ollama version is 0.9.6 [CODE=bash]ollama --version[/CODE] [HEADING=2]Linux[/HEADING] 4.1 Download & Install Ollama [CODE=bash]curl -fsSL https://ollama.com/install.sh | sh[/CODE] 4.2 Check Ollama version. ex ollama version is 0.9.6 [CODE=bash]ollama --version[/CODE] 5. Download model ඕන model එකක් මෙතනින් search කරගන්න. [URL]https://ollama.com/search[/URL] small model එකකින් පටන් ගමු. 5.1 gemma3:1b [URL]https://ollama.com/library/gemma3:1b[/URL] මේක lightweight censored model එකක්. single gpu එකක run කරගන්න පුලුවන්. 1B parameters තියෙනව. text input කරන්න පුලුවන්. 32k context window. 140 languages [CODE=bash]ollama pull gemma3:1b [/CODE] 5.2 gemma3:4b (recommended model) [URL]https://ollama.com/library/gemma3:4b[/URL] මේක lightweight censored multimodal එකක්. single gpu එකක run කරගන්න පුලුවන්. 4B parameters තියෙනව. image/text දෙකම input කරන්න පුලුවන්. 128k context window. 140 languages [CODE=bash]ollama pull gemma3:4b [/CODE] 6. List Downloaded Models [CODE=bash]ollama list[/CODE] 7. Run Your First Model [CODE=bash]ollama run gemma3:1b[/CODE] 8. To exit the interactive session, type /bye or press Ctrl + D 9. Install openweb ui 9.1 Create directory and navigate [CODE=bash] mkdir ~/open-webui-venv cd ~/open-webui-venv [/CODE] 9.2 Create a Python Virtual Environment [CODE=bash]python3 -m venv venv[/CODE] 9.3 Activate the Virtual Environment [CODE=bash]source venv/bin/activate[/CODE] You'll notice your terminal prompt changes (e.g., (venv) username@hostname:~/open-webui-venv$) to indicate that the virtual environment is active. 9.4 Install Open WebUI [CODE=bash]pip install open-webui[/CODE] 10.1 Start the Open WebUI Server [CODE=bash]open-webui serve[/CODE] This will usually start the server on [URL]http://localhost:8080[/URL]. You can then access it from your Windows browser [URL]http://localhost:8080[/URL] 10.2 Deactivate the Virtual Environment (when you're done) [CODE]deactivate[/CODE] [HEADING=3]Uncensored Qwen3 - JOSIEFIED 8B[/HEADING] [URL]https://ollama.com/goekdenizguelmez/JOSIEFIED-Qwen3[/URL] [CODE=bash]ollama run goekdenizguelmez/JOSIEFIED-Qwen3:8b[/CODE] This model has reduced safety filtering and may generate sensitive or controversial outputs. Use responsibly and at your own risk. [HEADING=3]Using Multimodal Models (LLaVA with Images)[/HEADING] [CODE=bash]ollama run llava "What do you see in this image? /home/yourusername/Pictures/my_image.jpg"[/CODE] (Replace /home/yourusername/Pictures/my_image.jpg with the actual path to your image.) මේ විදියට රන් කරගන්න පුලුවන් gpt, gemini වගේම models තියෙනව qwen3 කියන්නෙ එහෙම හොද කොලිටි මොඩල්ස්. 142GB 40K context, grok, deepseek r1, gemini 2.5, openai-o1 වලට වඩා benchmark score එක හොදයි. [URL]https://ollama.com/library/qwen3:235b[/URL] code කරන්න qwen coder එහෙම සුපිරි. OpenAI’s GPT-4oට වඩා හොදයි. [URL]https://ollama.com/library/qwen2.5-coder[/URL] ප්රශ්න මොනාහරි තියෙනව නම් දාපල්ල. gpu එකක් නැත්නම් gpu instance එකක රන් කරගන්න help එකක් ඕනෙ නම් කියපල්ල. [/QUOTE]
Insert quotes…
Verification
Dahaya deken beduwama keeyada?
Post reply
Top
Bottom