About r/LocalLLM
Subreddit to discuss locally run large language models.
The community at a glance
r/LocalLLM is a Subreddit for NLP Developers with roughly 194K members. It has been around since 2023. It uses a forum format for communication. On the Hive Index it ranks #8 in the NLP communities list.
Roughly 116K members have joined in the past year. Popular discussion topics include Model, Llm, and Local. Common discussion themes are Advice Requests and Solution Requests. Product recommendations often mention model, llm, and gpu.
Community Topics
Community Features
This community has a forum
Subreddit Analysis
via GummySearchMember growth over time
All time (yearly)
- 2024: 13K members
- 2025: 86K members
- 2026: 91K members
Past year (monthly)
- Sep: 4K members
- Oct: 3K members
- Nov: 6K members
- Dec: 5K members
- Jan: 7K members
- Feb: 10K members
- Mar: 11K members
- Apr: 17K members
- May: 17K members
- Jun: 16K members
- Jul: 14K members
Themes
- Advice Requests10 posts in the past month
- Local AI for game dev tooling, am I doing this wrong? (16GB & 24GB VRAM)
- Where do I start?
- Hardware list advice needed
#110Advice RequestsLocal AI for game dev tooling, am I doing this wrong? (16GB & 24GB VRAM) · Where do I start? · Hardware list advice needed - Solution Requests8 posts in the past month
- Is LM Studio safe to install on my personal mac? Any better (and safer) alternatives?
- Looking for a Local AI Tool for File & Folder Workflows (AMD RX 7600M XT, 8GB VRAM)
- Recommend for AI agent for booking platform and daily coding
#28Solution RequestsIs LM Studio safe to install on my personal mac? Any better (and safer) alternatives? · Looking for a Local AI Tool for File & Folder Workflows (AMD RX 7600M XT, 8GB VRAM) · Recommend for AI agent for booking platform and daily coding - Money Talk2 posts in the past month
- so $2.99 per 1 million tokens is expensive am I reading it right?
- Seeking tips for cutting costs on the ASUS Ascent GX10 AI Supercomputer.
#32Money Talkso $2.99 per 1 million tokens is expensive am I reading it right? · Seeking tips for cutting costs on the ASUS Ascent GX10 AI Supercomputer. - Ideas1 post in the past month
- Council 1.2: drop any AI's answer into a blind review by every other model you have
#41IdeasCouncil 1.2: drop any AI's answer into a blind review by every other model you have - Pain & Anger1 post in the past month
- Got Kimi K3 running on my MacBook. It's painfully slow, but it works.
#51Pain & AngerGot Kimi K3 running on my MacBook. It's painfully slow, but it works.
Topics
- Model39 posts in the past month
- Looking for a good copywriting model
- Is Colibri a taste of the future of local AI models?
- Has anyone else noticed Ollama Cloud usage changing for the same models over time?
#139ModelLooking for a good copywriting model · Is Colibri a taste of the future of local AI models? · Has anyone else noticed Ollama Cloud usage changing for the same models over time? - Llm24 posts in the past month#224Llm
- Local20 posts in the past month#320Local
- Ai18 posts in the past month#418Ai
- Looking For16 posts in the past month#516Looking For
Flair
- Question65 posts in the past month
- What am I missing? Self-Hosting Kimi K3 has 34× First-Year ROI at 90%
- so $2.99 per 1 million tokens is expensive am I reading it right?
- Tempted to upgrade to 6x RTX 3090 (144GB), is this setup future-proof for 120B models?
#165QuestionWhat am I missing? Self-Hosting Kimi K3 has 34× First-Year ROI at 90% · so $2.99 per 1 million tokens is expensive am I reading it right? · Tempted to upgrade to 6x RTX 3090 (144GB), is this setup future-proof for 120B models? - Discussion47 posts in the past month
- Thank you, whoever said don't quant the KV
- Got Kimi K3 running on my MacBook. It's painfully slow, but it works.
- Sir, nobody is supporting your bullshit regarding closed AI
#247DiscussionThank you, whoever said don't quant the KV · Got Kimi K3 running on my MacBook. It's painfully slow, but it works. · Sir, nobody is supporting your bullshit regarding closed AI - Project24 posts in the past month
- I built NightRun: boot a local LLM straight from a USB stick. No OS, just a UEFI app (x86-64 + Raspberry Pi 5)
- Ornith-397B running at Q4 on a single RTX PRO 6000 Blackwell 96GB - 2,354 tok/s prefill, ~20–24 tok/s decode
- I benchmarked 4 local models on an RX 7800 XT with contamination-proof tasks (seeded generation, no LLM judge) - there is no "best" model, only best-for-the-job
#324ProjectI built NightRun: boot a local LLM straight from a USB stick. No OS, just a UEFI app (x86-64 + Raspberry Pi 5) · Ornith-397B running at Q4 on a single RTX PRO 6000 Blackwell 96GB - 2,354 tok/s prefill, ~20–24 tok/s decode · I benchmarked 4 local models on an RX 7800 XT with contamination-proof tasks (seeded generation, no LLM judge) - there is no "best" model, only best-for-the-job - News10 posts in the past month
- When you're so desperate against open source you say stupid stuff that makes everyone and their mother come out to ridicule your opinion.
- Ran Moonshot's 2.8T-parameter Kimi K3 on a GPU-less mini-PC, one day after release
- You can now fine-tune my 3.96M-parameter TTS on your own voice or language
#410NewsWhen you're so desperate against open source you say stupid stuff that makes everyone and their mother come out to ridicule your opinion. · Ran Moonshot's 2.8T-parameter Kimi K3 on a GPU-less mini-PC, one day after release · You can now fine-tune my 3.96M-parameter TTS on your own voice or language - Research8 posts in the past month
- Amazing Performance from 2019 Mac Pro
- 96GB Ryzen AI 9 HX 370 on Minisforum N5 Pro as a daily-driver local LLM box: 13 models benchmarked, every flag, and everything I got wrong
- Lowest power consumption for iOS and MacOS on-device inference. LLMs, ASR, TTS. Apple SDK (iOS, macOS). Early access for developers!
#58ResearchAmazing Performance from 2019 Mac Pro · 96GB Ryzen AI 9 HX 370 on Minisforum N5 Pro as a daily-driver local LLM box: 13 models benchmarked, every flag, and everything I got wrong · Lowest power consumption for iOS and MacOS on-device inference. LLMs, ASR, TTS. Apple SDK (iOS, macOS). Early access for developers!
Product recommendations
- model16 posts in the past month
- Best models for 8x3090
- What the best model to run on m1 pro, 16gb ram for coders?
- best model for laptop and ram?
#116modelBest models for 8x3090 · What the best model to run on m1 pro, 16gb ram for coders? · best model for laptop and ram? - llm9 posts in the past month
- Best ultra low budget GPU for 70B and best LLM for my purpose
- Best LocalLLM for scientific theories and conversations?
- Best LLM to run locally on LM Studio (4GB VRAM) for extracting credit card statement PDFs into CSV/Excel?
#29llmBest ultra low budget GPU for 70B and best LLM for my purpose · Best LocalLLM for scientific theories and conversations? · Best LLM to run locally on LM Studio (4GB VRAM) for extracting credit card statement PDFs into CSV/Excel? - gpu4 posts in the past month
- Best ultra low budget GPU for 70B and best LLM for my purpose
- GPU recommendation for best possible LLM/AI/VR with 3000+€ budget
- Best Used Card For Running LLMS
#34gpuBest ultra low budget GPU for 70B and best LLM for my purpose · GPU recommendation for best possible LLM/AI/VR with 3000+€ budget · Best Used Card For Running LLMS
Community Reviews
Frequently asked questions
- Who is r/LocalLLM for?
- Best for NLP Developers enthusiasts looking for a Reddit-based community with forum discussion.
- Is r/LocalLLM free to join?
- This listing is not marked as paid-only. Access rules and any fees are decided by the community.
- How many members does r/LocalLLM have?
- Roughly 194K members, based on figures reported by the community or its host. Member counts are approximate and change over time.
- What platform is r/LocalLLM on?
- r/LocalLLM runs on Reddit. Reddit communities (or "subreddits") are forum-based groups on the popular social news aggregation, web content rating, and discussion website Reddit. Reddit is commonly known as "the front page of the internet". Users submit content to the site such as links, text posts, and images, which are then voted up or down and discussed by other members. From investing Reddit communities, to professional ones, to ones just for laughs, you're likely to find a community for you on Reddit.
- What topics does r/LocalLLM cover?
- On the Hive Index, r/LocalLLM is organized under NLP Developers.
- How do I join r/LocalLLM?
- You can join r/LocalLLM by clicking this link, or pressing the "Go to community" button above.
- What are the NLP communities like?
- Join NLP communities online to discuss and learn about the latest developments in language learning technology. You'll be able to chat with like-minded engineers and data scientists about the newest machine learning methods, from NLP to neural networks.
Monthly Stats
- 15
- Views /mo (+114%)
- 15
- Visitors /mo (+150%)
- 0
- Referrals /mo (-100%)
Similar Communities
7r/LanguageTechnology
This sub will focus on theory, careers, and applications of NLP (Natural Language Processing), which includes anything from Regex & Text Analytics to Transformers & LLMs. Language learning & copy/pasted ChatGPT conversations are outside the scope of the sub - please read the rules for more clarification.
OneAI
OneAI provides APIs to apply AI to text. Summarize conversations, categorize articles and detect user emotions. Join the AI dev community on discord!
Hugging Face Discuss
Official community for discussion of Hugging Face transformers, datasets, and NLP models
r/NLP
Neuro-Linguistic Programming (NLP) is an approach to communication, personal development, and psychotherapy created by Richard Bandler and John Grinder.
fast.ai Forum
Community for fast.ai courses, practical deep learning, and NLP applications
Kaggle NLP Community
Kaggle competitions and datasets for natural language processing