DEV Community

#llm

Posts

👋 Sign in for the ability to sort posts by relevant, latest, or top.
Designing MCP Tools for a 7B Model, Not a 70B One

Designing MCP Tools for a 7B Model, Not a 70B One

2
Comments 4
5 min read
LLMs on Consumer Hardware — Part 1: The Stack and First Benchmarks

LLMs on Consumer Hardware — Part 1: The Stack and First Benchmarks

Comments
3 min read
Medir si un LLM nombra a tu empresa: por qué una captura no sirve como métrica

Medir si un LLM nombra a tu empresa: por qué una captura no sirve como métrica

Comments
4 min read
The problem was never intelligence, it was continuity.

The problem was never intelligence, it was continuity.

Comments
2 min read
The LLM in my app is not allowed to decide anything

The LLM in my app is not allowed to decide anything

Comments 2
5 min read
How EvalPort's Grader System Works: 11 Types for LLM Evaluation

How EvalPort's Grader System Works: 11 Types for LLM Evaluation

Comments
2 min read
My Agent Orchestrator Burned 1-2M Opus Tokens Per Task. Here's the Postmortem.

My Agent Orchestrator Burned 1-2M Opus Tokens Per Task. Here's the Postmortem.

Comments 2
9 min read
The vendor documents this bug. A 30k-star repo shipped it anyway.

The vendor documents this bug. A 30k-star repo shipped it anyway.

Comments
12 min read
Identifying the Processor of a Bare-Metal Binary (Strategy 2): Testing LLMs

Identifying the Processor of a Bare-Metal Binary (Strategy 2): Testing LLMs

Comments
5 min read
Why LLMs Still Struggle With Tabular Prediction

Why LLMs Still Struggle With Tabular Prediction

Comments
4 min read
From Website URL to Useful AI Support Answers: A Practical Training Workflow

From Website URL to Useful AI Support Answers: A Practical Training Workflow

Comments
4 min read
DiffusionGemma Is Fast Because It Stops Pretending Text Has to Be Written Left to Right

DiffusionGemma Is Fast Because It Stops Pretending Text Has to Be Written Left to Right

2
Comments
3 min read
Spring AI Prompt Caching and Chat Memory: Where the Tokens Go — LLM Cost Control 2/4

Spring AI Prompt Caching and Chat Memory: Where the Tokens Go — LLM Cost Control 2/4

Comments 2
10 min read
MITRE ATLAS now has agentic attack techniques

MITRE ATLAS now has agentic attack techniques

1
Comments
4 min read
An LLM Attacked a Lab, Then Got Turned Around. That's Not the Scary Part.

An LLM Attacked a Lab, Then Got Turned Around. That's Not the Scary Part.

1
Comments
3 min read
👋 Sign in for the ability to sort posts by relevant, latest, or top.