Skip to content
Navigation menu
Search
Powered by Algolia
Search
Log in
Create account
DEV Community
Close
#
localllm
Follow
Hide
Posts
Left menu
đź‘‹
Sign in
for the ability to sort posts by
relevant
,
latest
, or
top
.
Right menu
Run Qwen3.8-Flash-Next Locally with EXL3 and TabbyAPI on 40+ GB of RAM
Seppe Gadeyne
Seppe Gadeyne
Seppe Gadeyne
Follow
Oct 4
Run Qwen3.8-Flash-Next Locally with EXL3 and TabbyAPI on 40+ GB of RAM
#
ai
#
localllm
#
llm
#
tutorial
Comments
Add Comment
10 min read
Qwen3.8-27B on one RTX 3090, and the setting its README says chat clients should leave at default
Reno Lu
Reno Lu
Reno Lu
Follow
Oct 2
Qwen3.8-27B on one RTX 3090, and the setting its README says chat clients should leave at default
#
vllm
#
localllm
#
speculativedecoding
#
quantization
Comments
Add Comment
4 min read
Still carrying the diagnosis
the kilted dev
the kilted dev
the kilted dev
Follow
Sep 29
Still carrying the diagnosis
#
localllm
#
debugging
#
buildinpublic
#
opensource
Comments
Add Comment
8 min read
Claude CLI 401 Unauthorized Refresh Token Issue
DevLog
DevLog
DevLog
Follow
Sep 28
Claude CLI 401 Unauthorized Refresh Token Issue
#
clitools
#
localllm
#
troubleshooting
Comments
Add Comment
3 min read
Discord bot not responding but still running: how to catch it
DevLog
DevLog
DevLog
Follow
Sep 27
Discord bot not responding but still running: how to catch it
#
automation
#
localllm
#
macos
Comments
Add Comment
3 min read
How Much RAM Do You Need for Local LLMs on a Mac?
Michael
Michael
Michael
Follow
Sep 23
How Much RAM Do You Need for Local LLMs on a Mac?
#
llm
#
localllm
#
ai
#
apple
Comments
Add Comment
3 min read
Claude Code MCP setup and usage rules
DevLog
DevLog
DevLog
Follow
Sep 23
Claude Code MCP setup and usage rules
#
clitools
#
macos
#
localllm
Comments
Add Comment
3 min read
Clone your own voice with a 17-second recording
DevLog
DevLog
DevLog
Follow
Sep 23
Clone your own voice with a 17-second recording
#
contentops
#
localllm
#
macos
Comments
Add Comment
2 min read
Prompt engineering: How to fix instructions that AI ignores
DevLog
DevLog
DevLog
Follow
Sep 21
Prompt engineering: How to fix instructions that AI ignores
#
promptengineering
#
ai
#
localllm
Comments
Add Comment
3 min read
Inside My llama.cpp Setup: Tuning Qwen 3.8 27B for 512K Context
Dmitry Amelchenko
Dmitry Amelchenko
Dmitry Amelchenko
Follow
Oct 3
Inside My llama.cpp Setup: Tuning Qwen 3.8 27B for 512K Context
#
ai
#
llm
#
localllm
#
llamacpp
Comments
2
 comments
7 min read
What Happens When You Ask an LLM a Question
Vishnu Hari Dadhich
Vishnu Hari Dadhich
Vishnu Hari Dadhich
Follow
Sep 14
What Happens When You Ask an LLM a Question
#
ai
#
localllm
#
llmbasics
#
gpu
Comments
Add Comment
8 min read
Run vLLM on Kubernetes with Minikube, WSL2 and NVIDIA GPU
Vishnu Hari Dadhich
Vishnu Hari Dadhich
Vishnu Hari Dadhich
Follow
Sep 14
Run vLLM on Kubernetes with Minikube, WSL2 and NVIDIA GPU
#
ai
#
localllm
#
vllm
#
kubernetes
Comments
Add Comment
13 min read
Testing the 9x Smaller Local LLM Claim on a GPU-less VPS
Qasim Parray
Qasim Parray
Qasim Parray
Follow
Sep 30
Testing the 9x Smaller Local LLM Claim on a GPU-less VPS
#
localllm
#
quantization
#
selfhosting
#
gguf
Comments
1
 comment
6 min read
I Ran DeepSeek V4 Flash Across Two DGX Sparks Over Ethernet
Kevin Tang
Kevin Tang
Kevin Tang
Follow
Sep 6
I Ran DeepSeek V4 Flash Across Two DGX Sparks Over Ethernet
#
ai
#
hardware
#
localllm
#
networking
Comments
Add Comment
11 min read
VRAM and RAM for local LLMs — honest planning bands, not a GPU tier list
Tech-Gurunomics
Tech-Gurunomics
Tech-Gurunomics
Follow
Sep 6
VRAM and RAM for local LLMs — honest planning bands, not a GPU tier list
#
localllm
#
llm
#
ollama
#
hardware
Comments
Add Comment
4 min read
đź‘‹
Sign in
for the ability to sort posts by
relevant
,
latest
, or
top
.
We're a place where coders share, stay up-to-date and grow their careers.
Log in
Create account