close

DEV Community

#benchmarking

Posts

đź‘‹ Sign in for the ability to sort posts by relevant, latest, or top.
SQLAlchemy ORM Security: The Raw Query Escape Hatch

SQLAlchemy ORM Security: The Raw Query Escape Hatch

Comments
5 min read
A Benchmark Smelled Funny

A Benchmark Smelled Funny

Comments 1
9 min read
If 30% of Coding Tasks May Be Broken, Your Leaderboard Needs an Uncertainty Budget

If 30% of Coding Tasks May Be Broken, Your Leaderboard Needs an Uncertainty Budget

BERJAYA 1
Comments
3 min read
Benchmarking Apple's SpeechAnalyzer API vs. Whisper: Performance, Accuracy, and Use Cases

Benchmarking Apple's SpeechAnalyzer API vs. Whisper: Performance, Accuracy, and Use Cases

Comments
2 min read
IdeaGene-Bench: A New Benchmark for Scientific Lineage Reasoning in AI

IdeaGene-Bench: A New Benchmark for Scientific Lineage Reasoning in AI

Comments
4 min read
How I Benchmarked an LLM Running Entirely on a Phone (No Cloud, No API)

How I Benchmarked an LLM Running Entirely on a Phone (No Cloud, No API)

Comments
16 min read
prima.cpp local llm benchmark: 15% Faster Than llama.cpp

prima.cpp local llm benchmark: 15% Faster Than llama.cpp

Comments
8 min read
My Code, My Test, and My Prompt All Agreed. All Three Were Wrong.

My Code, My Test, and My Prompt All Agreed. All Three Were Wrong.

Comments
10 min read
Building an Official Performance Baseline for Vix.cpp Core v2.6.3

Building an Official Performance Baseline for Vix.cpp Core v2.6.3

Comments
3 min read
I measure how fast 42 LLMs actually answer. Here's the honest method.

I measure how fast 42 LLMs actually answer. Here's the honest method.

BERJAYA 1
Comments 1
2 min read
Comparing Node.js Postgres Client Libraries: brianc/node-postgres vs. porsager/postgres for Efficiency and Use Cases

Comparing Node.js Postgres Client Libraries: brianc/node-postgres vs. porsager/postgres for Efficiency and Use Cases

BERJAYA 1
Comments
10 min read
I benchmarked my document-extraction API against Textract and Google DocAI — on public datasets, in public CI

I benchmarked my document-extraction API against Textract and Google DocAI — on public datasets, in public CI

Comments 1
3 min read
My AI memory benchmark said 98.3%. The number was true — and worthless.

My AI memory benchmark said 98.3%. The number was true — and worthless.

Comments 14
4 min read
The Mean Is Lying to You: Benchmarks Hide the Variance That Breaks Prod

The Mean Is Lying to You: Benchmarks Hide the Variance That Breaks Prod

BERJAYA 1
Comments
5 min read
Is AI-Generated Code Buggier? The 2025-26 Data

Is AI-Generated Code Buggier? The 2025-26 Data

Comments 1
3 min read
đź‘‹ Sign in for the ability to sort posts by relevant, latest, or top.