DEV Community

#llm

Posts

👋 Sign in for the ability to sort posts by relevant, latest, or top.
One TPU Chip, Eight Agents: Serving Small Agent Workloads with Raw JAX

One TPU Chip, Eight Agents: Serving Small Agent Workloads with Raw JAX

3
Comments 2
15 min read
Testing Non-Deterministic LLM Pipelines in CI: A Contract-Based Approach

Testing Non-Deterministic LLM Pipelines in CI: A Contract-Based Approach

2
Comments 1
4 min read
I gave the same fabricated answer to RAGAS and DeepEval. One scored it 0.0. The other scored it 1.0

I gave the same fabricated answer to RAGAS and DeepEval. One scored it 0.0. The other scored it 1.0

Comments
6 min read
Your cache_read_input_tokens is zero. Here is what silently did it.

Your cache_read_input_tokens is zero. Here is what silently did it.

Comments
5 min read
Close your editor before heavy jobs? The heavy job lives inside my editor

Close your editor before heavy jobs? The heavy job lives inside my editor

Comments
4 min read
Four Models Cited My Numbers Perfectly. One Still Misread Them.

Four Models Cited My Numbers Perfectly. One Still Misread Them.

Comments
6 min read
OpenEval: Why LLM Evaluation Needs a Standard Format

OpenEval: Why LLM Evaluation Needs a Standard Format

Comments
1 min read
Your AI Subagents Are Lying to You: 4 Silent Failure Modes

Your AI Subagents Are Lying to You: 4 Silent Failure Modes

1
Comments 3
4 min read
Why does parsing scientific papers for RAG still break on equations and tables?

Why does parsing scientific papers for RAG still break on equations and tables?

2
Comments
3 min read
Building a unified OpenAI-compatible gateway for 16 Chinese LLMs: architecture notes on routing, payments and ops

Building a unified OpenAI-compatible gateway for 16 Chinese LLMs: architecture notes on routing, payments and ops

Comments 1
3 min read
`finish_reason=length` Returned Empty Content — and the Error Message Lied to Me

`finish_reason=length` Returned Empty Content — and the Error Message Lied to Me

1
Comments
3 min read
Model + Harness = Agent: The Gap Isn’t Where You Think

Model + Harness = Agent: The Gap Isn’t Where You Think

Comments
4 min read
I Trust My AI Completely—Except When It Says “Done”

I Trust My AI Completely—Except When It Says “Done”

1
Comments 1
4 min read
How to Build an AI Kill Switch (and Why Every Agent Needs One)

How to Build an AI Kill Switch (and Why Every Agent Needs One)

1
Comments
9 min read
From RAG to Agentic AI. How I Added LangGraph to My Local

From RAG to Agentic AI. How I Added LangGraph to My Local

1
Comments 1
6 min read
👋 Sign in for the ability to sort posts by relevant, latest, or top.