Skip to content
Navigation menu
Search
Powered by Algolia
Search
Log in
Create account
DEV Community
Close
#
localllm
Follow
Hide
Posts
Left menu
👋
Sign in
for the ability to sort posts by
relevant
,
latest
, or
top
.
Right menu
I Ran DeepSeek V4 Flash Across Two DGX Sparks Over Ethernet
Kevin Tang
Kevin Tang
Kevin Tang
Follow
Sep 6
I Ran DeepSeek V4 Flash Across Two DGX Sparks Over Ethernet
#
ai
#
hardware
#
localllm
#
networking
Comments
Add Comment
11 min read
VRAM and RAM for local LLMs — honest planning bands, not a GPU tier list
Tech-Gurunomics
Tech-Gurunomics
Tech-Gurunomics
Follow
Sep 6
VRAM and RAM for local LLMs — honest planning bands, not a GPU tier list
#
localllm
#
llm
#
ollama
#
hardware
Comments
Add Comment
4 min read
Running Ollama on a 32 GB MacBook Air: A Practical First Setup
Vishnu Hari Dadhich
Vishnu Hari Dadhich
Vishnu Hari Dadhich
Follow
Sep 7
Running Ollama on a 32 GB MacBook Air: A Practical First Setup
#
ai
#
localllm
#
ollama
#
applesilicon
Comments
Add Comment
6 min read
Running llama.cpp on a 32 GB MacBook Air: A Direct Comparison with Ollama
Vishnu Hari Dadhich
Vishnu Hari Dadhich
Vishnu Hari Dadhich
Follow
Sep 7
Running llama.cpp on a 32 GB MacBook Air: A Direct Comparison with Ollama
#
ai
#
localllm
#
llamacpp
#
ollama
Comments
1
 comment
9 min read
Running a 35B MoE Model on an 8 GB Laptop GPU: Testing FreeToken
Vishnu Hari Dadhich
Vishnu Hari Dadhich
Vishnu Hari Dadhich
Follow
Sep 7
Running a 35B MoE Model on an 8 GB Laptop GPU: Testing FreeToken
#
ai
#
localllm
#
freetoken
#
gpu
Comments
3
 comments
7 min read
Temperature 0 is not reproducible. I measured 30 percent of my output changing between identical runs.
Rickesh T N
Rickesh T N
Rickesh T N
Follow
Sep 5
Temperature 0 is not reproducible. I measured 30 percent of my output changing between identical runs.
#
llm
#
benchmarks
#
localllm
#
reproducibility
1
 reaction
Comments
1
 comment
3 min read
Our 4B beat Claude Opus on a 440K-token corpus. Then it came last on the public benchmark.
Rickesh T N
Rickesh T N
Rickesh T N
Follow
Sep 5
Our 4B beat Claude Opus on a 440K-token corpus. Then it came last on the public benchmark.
#
llm
#
benchmarks
#
localllm
#
privateai
2
 reactions
Comments
Add Comment
4 min read
I told the model to separate fields with <TAB>. It did exactly that, and I lost 79 percent of my data.
Rickesh T N
Rickesh T N
Rickesh T N
Follow
Sep 5
I told the model to separate fields with <TAB>. It did exactly that, and I lost 79 percent of my data.
#
llm
#
prompting
#
debugging
#
localllm
1
 reaction
Comments
Add Comment
3 min read
A 4B on a 6GB laptop matched frontier-model accuracy on aggregation — except when the answer is a number
Rickesh T N
Rickesh T N
Rickesh T N
Follow
Sep 5
A 4B on a 6GB laptop matched frontier-model accuracy on aggregation — except when the answer is a number
#
llm
#
benchmarks
#
localllm
#
agents
1
 reaction
Comments
Add Comment
4 min read
Your agent truncates the corpus and answers anyway. Two harnesses, and a router that picks between them.
Rickesh T N
Rickesh T N
Rickesh T N
Follow
Sep 5
Your agent truncates the corpus and answers anyway. Two harnesses, and a router that picks between them.
#
llm
#
agents
#
localllm
#
benchmarks
1
 reaction
Comments
Add Comment
5 min read
How to Run a Free AI Coding Assistant Locally with VS Code, opencode, and LM Studio
Aravinda åŠ é˜³
Aravinda åŠ é˜³
Aravinda åŠ é˜³
Follow
Sep 6
How to Run a Free AI Coding Assistant Locally with VS Code, opencode, and LM Studio
#
tutorials
#
localllm
#
aicodingassistant
#
lmstudio
Comments
Add Comment
4 min read
What really fits in 8GB VRAM
the kilted dev
the kilted dev
the kilted dev
Follow
Aug 18
What really fits in 8GB VRAM
#
localllm
#
vram
#
gpu
#
buildinpublic
Comments
Add Comment
7 min read
Moving Scheduled LLM Curation from Cloud APIs to Local Models
Guatu
Guatu
Guatu
Follow
Aug 14
Moving Scheduled LLM Curation from Cloud APIs to Local Models
#
aiagents
#
localllm
#
ollama
#
kubernetes
Comments
Add Comment
9 min read
Nine ways to talk to a local model
the kilted dev
the kilted dev
the kilted dev
Follow
Aug 11
Nine ways to talk to a local model
#
localllm
#
ollama
#
llamacpp
#
buildinpublic
Comments
Add Comment
9 min read
A 4B model on a 6GB laptop beat Claude Opus on our 440K-token corpus. The fix was giving the model less to do.
Rickesh T N
Rickesh T N
Rickesh T N
Follow
Aug 24
A 4B model on a 6GB laptop beat Claude Opus on our 440K-token corpus. The fix was giving the model less to do.
#
privateai
#
llm
#
inference
#
localllm
1
 reaction
Comments
1
 comment
4 min read
👋
Sign in
for the ability to sort posts by
relevant
,
latest
, or
top
.
We're a place where coders share, stay up-to-date and grow their careers.
Log in
Create account