DEV Community

#llm

Posts

👋 Sign in for the ability to sort posts by relevant, latest, or top.
How do you stop an LLM from leaking API keys in the code it writes? Default to secret

How do you stop an LLM from leaking API keys in the code it writes? Default to secret

6
Comments 4
8 min read
RAG ranking is not the same as judging with Jev

RAG ranking is not the same as judging with Jev

1
Comments
9 min read
I Cut 2,490 Agent Test Runs to 206 and Kept the Same Coverage

I Cut 2,490 Agent Test Runs to 206 and Kept the Same Coverage

7
Comments 1
5 min read
Your AI eval is green because it never called the model

Your AI eval is green because it never called the model

Comments 4
6 min read
decider: one forward pass, typed decisions, calibrated probabilities

decider: one forward pass, typed decisions, calibrated probabilities

Comments
5 min read
Memory Is a System, Not a Prompt: Putting the Stack to Work

Memory Is a System, Not a Prompt: Putting the Stack to Work

1
Comments
5 min read
Think smaller: why specialist SLMs beat frontier models in production

Think smaller: why specialist SLMs beat frontier models in production

Comments
8 min read
Resisting Mode Gravity: Why Bigger LLMs Produce Mediocre Output

Resisting Mode Gravity: Why Bigger LLMs Produce Mediocre Output

Comments
9 min read
I gave my local AI agent background workers. Then it tried to deploy to production.

I gave my local AI agent background workers. Then it tried to deploy to production.

2
Comments 3
6 min read
Designing an eval harness for prompt-injection detection: what measuring my defenses actually taught me

Designing an eval harness for prompt-injection detection: what measuring my defenses actually taught me

Comments
6 min read
Jev is now open to everyone: what a "System One" model costs, and how to start with $5 in free credit

Jev is now open to everyone: what a "System One" model costs, and how to start with $5 in free credit

Comments
1 min read
8 of my AI agent's 30 test calls failed. Every one was my fault.

8 of my AI agent's 30 test calls failed. Every one was my fault.

Comments 1
7 min read
Fine-Tune, Deploy and Use LLM As AI Agent

Fine-Tune, Deploy and Use LLM As AI Agent

Comments
3 min read
Unlocking Client-Side AI: Running LLMs in the Browser with WebGPU

Unlocking Client-Side AI: Running LLMs in the Browser with WebGPU

Comments
7 min read
Treat streamed LLM output as an uncommitted draft

Treat streamed LLM output as an uncommitted draft

Comments
4 min read
👋 Sign in for the ability to sort posts by relevant, latest, or top.