DEV Community

#llm

Posts

👋 Sign in for the ability to sort posts by relevant, latest, or top.
Mistral Small 3.2 Lands With Sharper Function Calling and a 128K Context Window

Mistral Small 3.2 Lands With Sharper Function Calling and a 128K Context Window

1
Comments
4 min read
Demystifying LLM Context Windows: How AI Memory Works (and Why It Fails)

Demystifying LLM Context Windows: How AI Memory Works (and Why It Fails)

Comments
7 min read
I let my own 31B model take over development of the thing running it

I let my own 31B model take over development of the thing running it

Comments
3 min read
GPT-6 Astra, Claude Fable, Gemini 3.8: A Busy Week for New LLM Releases

GPT-6 Astra, Claude Fable, Gemini 3.8: A Busy Week for New LLM Releases

Comments
3 min read
Seven months of self-hosting our own AI stack: four bugs I can point at in the changelog

Seven months of self-hosting our own AI stack: four bugs I can point at in the changelog

Comments
5 min read
Installing GPT4All, an Open-Source Chatbot Application for Running LLMs

Installing GPT4All, an Open-Source Chatbot Application for Running LLMs

10
Picked as gem Comments
9 min read
From API to GPU, Week 6 (Part 2): Watching a Neural Network Learn

From API to GPU, Week 6 (Part 2): Watching a Neural Network Learn

Comments
17 min read
From API to GPU, Week 6 (Part 1): A Model That Predicts, and How Wrong It Is

From API to GPU, Week 6 (Part 1): A Model That Predicts, and How Wrong It Is

Comments
13 min read
Un gateway para saber qué producto me quema la factura de LLM (y el día que mi README mintió)

Un gateway para saber qué producto me quema la factura de LLM (y el día que mi README mintió)

Comments
3 min read
Fable's Fumble into a Touchdown

Fable's Fumble into a Touchdown

Comments
5 min read
When Confidence Lies: Engineering Uncertainty-Aware AI Control Loops for High-Stakes Production Systems

When Confidence Lies: Engineering Uncertainty-Aware AI Control Loops for High-Stakes Production Systems

Comments 1
7 min read
Fine-tuning a 1.7B model at 3.2 GB VRAM — building FineTune Studio

Fine-tuning a 1.7B model at 3.2 GB VRAM — building FineTune Studio

Comments
2 min read
Your Local LLM Has a Hidden Context Limit

Your Local LLM Has a Hidden Context Limit

Comments
4 min read
How We Built Perceive: Web Content Extraction for RAG Pipelines

How We Built Perceive: Web Content Extraction for RAG Pipelines

Comments
8 min read
Renting GPUs for Crypto: I Did the Math on Self-Hosting GLM-5.2

Renting GPUs for Crypto: I Did the Math on Self-Hosting GLM-5.2

Comments
6 min read
👋 Sign in for the ability to sort posts by relevant, latest, or top.