close

DEV Community

#llm

Posts

👋 Sign in for the ability to sort posts by relevant, latest, or top.
Just Train More: Measuring the Exchange Rate

Just Train More: Measuring the Exchange Rate

Comments 1
5 min read
Why Local LLMs Don't Need C++ or Python: Building a 15MB Native AOT Inference Engine in .NET 10

Why Local LLMs Don't Need C++ or Python: Building a 15MB Native AOT Inference Engine in .NET 10

BERJAYA 1
Comments 5
5 min read
Your LLM cost estimate is wrong above 200,000 tokens

Your LLM cost estimate is wrong above 200,000 tokens

Comments
2 min read
Nvidia เปิดสูตรเหรียญทอง IMO ทั้งชุด หลังนักคณิตศาสตร์เตือนเรื่อง AI

Nvidia เปิดสูตรเหรียญทอง IMO ทั้งชุด หลังนักคณิตศาสตร์เตือนเรื่อง AI

Comments
1 min read
AGI Definition, So You Don't Get Tricked Anymore

AGI Definition, So You Don't Get Tricked Anymore

BERJAYA 1
Comments
3 min read
The Meter Is Running on Every Request

The Meter Is Running on Every Request

BERJAYA 1
Comments
2 min read
Run Claude Code on MiniMax-M3 with a 1M context for about $0.30 per million input tokens

Run Claude Code on MiniMax-M3 with a 1M context for about $0.30 per million input tokens

Comments
5 min read
Claude Fable 5.1: o que muda de verdade no topo da linha da Anthropic

Claude Fable 5.1: o que muda de verdade no topo da linha da Anthropic

BERJAYA 1
Comments
3 min read
LLM Inference Optimization: Techniques for Faster and Cheaper AI

LLM Inference Optimization: Techniques for Faster and Cheaper AI

Comments
2 min read
Let the LLM be the author, not a scripted text formatter

Let the LLM be the author, not a scripted text formatter

Comments
1 min read
Agent-Cache: Multi-Tier LLM Caching for Valkey and Redis

Agent-Cache: Multi-Tier LLM Caching for Valkey and Redis

BERJAYA 1
Comments
6 min read
AX-RAY & K-MYTHOS: Inside Korea's Consortium-Built Security-Specialized AI Foundation Model

AX-RAY & K-MYTHOS: Inside Korea's Consortium-Built Security-Specialized AI Foundation Model

BERJAYA 1
Comments
4 min read
Spec-Driven Development (SDD): De la improvisación a la ingeniería con agentes de IA

Spec-Driven Development (SDD): De la improvisación a la ingeniería con agentes de IA

Comments
8 min read
A RAG chatbot on your company knowledge base: what it is and when it pays off

A RAG chatbot on your company knowledge base: what it is and when it pays off

BERJAYA 1
Comments
5 min read
Practice RAG Retrieval Metrics Offline — A Tiny Stdlib Eval Loop (Synthetic Data)

Practice RAG Retrieval Metrics Offline — A Tiny Stdlib Eval Loop (Synthetic Data)

Comments 2
5 min read
👋 Sign in for the ability to sort posts by relevant, latest, or top.