DEV Community

#llm

Posts

👋 Sign in for the ability to sort posts by relevant, latest, or top.
We Could Have Shipped on Local Models Alone

We Could Have Shipped on Local Models Alone

5
Comments
12 min read
Four versions of one translate() function

Four versions of one translate() function

Comments
3 min read
Giving AI Agents the Same RBAC Rules as Your Users: Building a Laravel Permission Layer LLMs Actually Respect

Giving AI Agents the Same RBAC Rules as Your Users: Building a Laravel Permission Layer LLMs Actually Respect

5
Comments
11 min read
Make Your Code Review Agent Write Down How the Bug Actually Happens

Make Your Code Review Agent Write Down How the Bug Actually Happens

2
Comments 3
6 min read
I built a prompt injection detection API that responds in <1ms — here's how

I built a prompt injection detection API that responds in <1ms — here's how

Comments
2 min read
Small language models (1B–3B): what they're good for after fine-tuning

Small language models (1B–3B): what they're good for after fine-tuning

Comments
3 min read
Instruction tuning vs domain adaptation: two different fine-tuning goals

Instruction tuning vs domain adaptation: two different fine-tuning goals

Comments
3 min read
QLoRA hyperparameters that actually matter (rank, alpha, learning rate, epochs)

QLoRA hyperparameters that actually matter (rank, alpha, learning rate, epochs)

Comments
3 min read
With Context Windows This Large, Why Do We Still Need Memory?

With Context Windows This Large, Why Do We Still Need Memory?

Comments
5 min read
LoRA vs QLoRA vs full fine-tuning: cost, quality and when each makes sense

LoRA vs QLoRA vs full fine-tuning: cost, quality and when each makes sense

Comments
3 min read
Inside nano-vLLM: What an RTX 3090 Reveals About LLM Serving

Inside nano-vLLM: What an RTX 3090 Reveals About LLM Serving

Comments
14 min read
It Fit in Memory and Was Still Unusable — Do the Bandwidth Arithmetic First

It Fit in Memory and Was Still Unusable — Do the Bandwidth Arithmetic First

2
Comments 1
5 min read
VRAM and RAM for local LLMs — honest planning bands, not a GPU tier list

VRAM and RAM for local LLMs — honest planning bands, not a GPU tier list

Comments
4 min read
The Local LLM Weight Classes, September 2026: What Actually Fits on Your Machine

The Local LLM Weight Classes, September 2026: What Actually Fits on Your Machine

Comments
9 min read
🧠 AI Context Engineering (Part 5): Context Optimization - Give AI What It Needs, Not Everything You Have

🧠 AI Context Engineering (Part 5): Context Optimization - Give AI What It Needs, Not Everything You Have

1
Comments
11 min read
👋 Sign in for the ability to sort posts by relevant, latest, or top.