r/LocalLLM

Community Overview

About r/LocalLLM

Subreddit to discuss locally run large language models.

The community at a glance

r/LocalLLM is a Subreddit for NLP Developers with roughly 194K members. It has been around since 2023. It uses a forum format for communication. On the Hive Index it ranks #8 in the NLP communities list.

Roughly 116K members have joined in the past year. Popular discussion topics include Model, Llm, and Local. Common discussion themes are Advice Requests and Solution Requests. Product recommendations often mention model, llm, and gpu.

On Reddit
Established 2023
194K Members

Community Features

This community has a forum

Subreddit Analysis

via GummySearch
Yearly: +116K members
Growth: +150.3% / year

Member growth over time

All time (yearly)

  • 2024: 13K members
  • 2025: 86K members
  • 2026: 91K members

Past year (monthly)

  • Sep: 4K members
  • Oct: 3K members
  • Nov: 6K members
  • Dec: 5K members
  • Jan: 7K members
  • Feb: 10K members
  • Mar: 11K members
  • Apr: 17K members
  • May: 17K members
  • Jun: 16K members
  • Jul: 14K members

Themes

  • Advice Requests
    10 posts in the past month
    1. Local AI for game dev tooling, am I doing this wrong? (16GB & 24GB VRAM)
    2. Where do I start?
    3. Hardware list advice needed
    #1
    Advice Requests
    Local AI for game dev tooling, am I doing this wrong? (16GB & 24GB VRAM) · Where do I start? · Hardware list advice needed
    10
  • Solution Requests
    8 posts in the past month
    1. Is LM Studio safe to install on my personal mac? Any better (and safer) alternatives?
    2. Looking for a Local AI Tool for File & Folder Workflows (AMD RX 7600M XT, 8GB VRAM)
    3. Recommend for AI agent for booking platform and daily coding
    #2
    Solution Requests
    Is LM Studio safe to install on my personal mac? Any better (and safer) alternatives? · Looking for a Local AI Tool for File & Folder Workflows (AMD RX 7600M XT, 8GB VRAM) · Recommend for AI agent for booking platform and daily coding
    8
  • Money Talk
    2 posts in the past month
    1. so $2.99 per 1 million tokens is expensive am I reading it right?
    2. Seeking tips for cutting costs on the ASUS Ascent GX10 AI Supercomputer.
    #3
    Money Talk
    so $2.99 per 1 million tokens is expensive am I reading it right? · Seeking tips for cutting costs on the ASUS Ascent GX10 AI Supercomputer.
    2
  • Ideas
    1 post in the past month
    1. Council 1.2: drop any AI's answer into a blind review by every other model you have
    #4
    Ideas
    Council 1.2: drop any AI's answer into a blind review by every other model you have
    1
  • Pain & Anger
    1 post in the past month
    1. Got Kimi K3 running on my MacBook. It's painfully slow, but it works.
    #5
    Pain & Anger
    Got Kimi K3 running on my MacBook. It's painfully slow, but it works.
    1

Topics

  • Model
    39 posts in the past month
    1. Looking for a good copywriting model
    2. Is Colibri a taste of the future of local AI models?
    3. Has anyone else noticed Ollama Cloud usage changing for the same models over time?
    #1
    Model
    Looking for a good copywriting model · Is Colibri a taste of the future of local AI models? · Has anyone else noticed Ollama Cloud usage changing for the same models over time?
    39
  • Llm
    24 posts in the past month
    #2
    Llm
    24
  • Local
    20 posts in the past month
    #3
    Local
    20
  • Ai
    18 posts in the past month
    #4
    Ai
    18
  • Looking For
    16 posts in the past month
    #5
    Looking For
    16

Flair

  • Question
    65 posts in the past month
    1. What am I missing? Self-Hosting Kimi K3 has 34× First-Year ROI at 90%
    2. so $2.99 per 1 million tokens is expensive am I reading it right?
    3. Tempted to upgrade to 6x RTX 3090 (144GB), is this setup future-proof for 120B models?
    #1
    Question
    What am I missing? Self-Hosting Kimi K3 has 34× First-Year ROI at 90% · so $2.99 per 1 million tokens is expensive am I reading it right? · Tempted to upgrade to 6x RTX 3090 (144GB), is this setup future-proof for 120B models?
    65
  • Discussion
    47 posts in the past month
    1. Thank you, whoever said don't quant the KV
    2. Got Kimi K3 running on my MacBook. It's painfully slow, but it works.
    3. Sir, nobody is supporting your bullshit regarding closed AI
    #2
    Discussion
    Thank you, whoever said don't quant the KV · Got Kimi K3 running on my MacBook. It's painfully slow, but it works. · Sir, nobody is supporting your bullshit regarding closed AI
    47
  • Project
    24 posts in the past month
    1. I built NightRun: boot a local LLM straight from a USB stick. No OS, just a UEFI app (x86-64 + Raspberry Pi 5)
    2. Ornith-397B running at Q4 on a single RTX PRO 6000 Blackwell 96GB - 2,354 tok/s prefill, ~20–24 tok/s decode
    3. I benchmarked 4 local models on an RX 7800 XT with contamination-proof tasks (seeded generation, no LLM judge) - there is no "best" model, only best-for-the-job
    #3
    Project
    I built NightRun: boot a local LLM straight from a USB stick. No OS, just a UEFI app (x86-64 + Raspberry Pi 5) · Ornith-397B running at Q4 on a single RTX PRO 6000 Blackwell 96GB - 2,354 tok/s prefill, ~20–24 tok/s decode · I benchmarked 4 local models on an RX 7800 XT with contamination-proof tasks (seeded generation, no LLM judge) - there is no "best" model, only best-for-the-job
    24
  • News
    10 posts in the past month
    1. When you're so desperate against open source you say stupid stuff that makes everyone and their mother come out to ridicule your opinion.
    2. Ran Moonshot's 2.8T-parameter Kimi K3 on a GPU-less mini-PC, one day after release
    3. You can now fine-tune my 3.96M-parameter TTS on your own voice or language
    #4
    News
    When you're so desperate against open source you say stupid stuff that makes everyone and their mother come out to ridicule your opinion. · Ran Moonshot's 2.8T-parameter Kimi K3 on a GPU-less mini-PC, one day after release · You can now fine-tune my 3.96M-parameter TTS on your own voice or language
    10
  • Research
    8 posts in the past month
    1. Amazing Performance from 2019 Mac Pro
    2. 96GB Ryzen AI 9 HX 370 on Minisforum N5 Pro as a daily-driver local LLM box: 13 models benchmarked, every flag, and everything I got wrong
    3. Lowest power consumption for iOS and MacOS on-device inference. LLMs, ASR, TTS. Apple SDK (iOS, macOS). Early access for developers!
    #5
    Research
    Amazing Performance from 2019 Mac Pro · 96GB Ryzen AI 9 HX 370 on Minisforum N5 Pro as a daily-driver local LLM box: 13 models benchmarked, every flag, and everything I got wrong · Lowest power consumption for iOS and MacOS on-device inference. LLMs, ASR, TTS. Apple SDK (iOS, macOS). Early access for developers!
    8

Product recommendations

  • model
    16 posts in the past month
    1. Best models for 8x3090
    2. What the best model to run on m1 pro, 16gb ram for coders?
    3. best model for laptop and ram?
    #1
    model
    Best models for 8x3090 · What the best model to run on m1 pro, 16gb ram for coders? · best model for laptop and ram?
    16
  • llm
    9 posts in the past month
    1. Best ultra low budget GPU for 70B and best LLM for my purpose
    2. Best LocalLLM for scientific theories and conversations?
    3. Best LLM to run locally on LM Studio (4GB VRAM) for extracting credit card statement PDFs into CSV/Excel?
    #2
    llm
    Best ultra low budget GPU for 70B and best LLM for my purpose · Best LocalLLM for scientific theories and conversations? · Best LLM to run locally on LM Studio (4GB VRAM) for extracting credit card statement PDFs into CSV/Excel?
    9
  • gpu
    4 posts in the past month
    1. Best ultra low budget GPU for 70B and best LLM for my purpose
    2. GPU recommendation for best possible LLM/AI/VR with 3000+€ budget
    3. Best Used Card For Running LLMS
    #3
    gpu
    Best ultra low budget GPU for 70B and best LLM for my purpose · GPU recommendation for best possible LLM/AI/VR with 3000+€ budget · Best Used Card For Running LLMS
    4

Frequently asked questions

Who is r/LocalLLM for?
Best for NLP Developers enthusiasts looking for a Reddit-based community with forum discussion.
Is r/LocalLLM free to join?
This listing is not marked as paid-only. Access rules and any fees are decided by the community.
How many members does r/LocalLLM have?
Roughly 194K members, based on figures reported by the community or its host. Member counts are approximate and change over time.
What platform is r/LocalLLM on?
r/LocalLLM runs on Reddit. Reddit communities (or "subreddits") are forum-based groups on the popular social news aggregation, web content rating, and discussion website Reddit. Reddit is commonly known as "the front page of the internet". Users submit content to the site such as links, text posts, and images, which are then voted up or down and discussed by other members. From investing Reddit communities, to professional ones, to ones just for laughs, you're likely to find a community for you on Reddit.
What topics does r/LocalLLM cover?
On the Hive Index, r/LocalLLM is organized under NLP Developers.
How do I join r/LocalLLM?
You can join r/LocalLLM by clicking this link, or pressing the "Go to community" button above.
What are the NLP communities like?
Join NLP communities online to discuss and learn about the latest developments in language learning technology. You'll be able to chat with like-minded engineers and data scientists about the newest machine learning methods, from NLP to neural networks.