close

DEV Community

plasma profile picture

plasma

Building TokenBay, an OpenAl-compatible API gateway for multiple Al models.

Location silicon valley Joined Joined on  Personal website https://www.tokenbay.com/
I Couldn’t Fix My LLM Costs Until I Measured Tokens Per Feature

I Couldn’t Fix My LLM Costs Until I Measured Tokens Per Feature

Comments
7 min read

Want to connect with plasma?

Create an account to connect with plasma. You can also sign in below to proceed if you already have an account.

Already have an account? Sign in
The LLM Response Looked Fine. My Parser Disagreed.

The LLM Response Looked Fine. My Parser Disagreed.

BERJAYA 1
Comments
4 min read
A Tiny LLM Request Recorder I Use to Reproduce Production Failures

A Tiny LLM Request Recorder I Use to Reproduce Production Failures

BERJAYA 1
Comments 1
5 min read
My LLM Provider Returned 200. The Workflow Still Failed.

My LLM Provider Returned 200. The Workflow Still Failed.

Comments
6 min read
My Agent Said “Done” After One Tool Call Failed Silently

My Agent Said “Done” After One Tool Call Failed Silently

Comments
6 min read
My LLM Bill Kept Growing, but User Traffic Didn’t

My LLM Bill Kept Growing, but User Traffic Didn’t

BERJAYA 3
Comments 5
6 min read
I Stopped Swapping LLM Providers Without a Smoke Test

I Stopped Swapping LLM Providers Without a Smoke Test

Comments
6 min read
A Small Node.js Wrapper for LLM API Retries, Timeouts, and Logging

A Small Node.js Wrapper for LLM API Retries, Timeouts, and Logging

BERJAYA 2
Comments 4
5 min read
The LLM API Failure Policy I Wish I Had Before My First Production Incident

Branching logic for RPM vs TPM rate limits

The LLM API Failure Policy I Wish I Had Before My First Production Incident

BERJAYA BERJAYA BERJAYA 6
Comments 15
6 min read
The Retry Setup I Use for LLM APIs Without Accidentally Duplicating User Actions

The Retry Setup I Use for LLM APIs Without Accidentally Duplicating User Actions

Comments
7 min read
Stop Treating LLM API Errors Like Normal HTTP Errors

Stop Treating LLM API Errors Like Normal HTTP Errors

Comments 6
7 min read
LLM API debugging checklist

LLM API debugging checklist

Comments
7 min read
What I Log When an LLM API Call Fails Mid-Stream

What I Log When an LLM API Call Fails Mid-Stream

Comments
8 min read
The LLM API Timeout Playbook I Wish I Had Before Production

The LLM API Timeout Playbook I Wish I Had Before Production

Comments
8 min read
My LLM API Calls Were Failing Silently. Here's the Logging Setup I Wish I Had Earlier

My LLM API Calls Were Failing Silently. Here's the Logging Setup I Wish I Had Earlier

BERJAYA 3
Comments 4
8 min read
OpenAI-Compatible APIs Are Great Until Streaming Breaks: What I Check Before Switching Providers

OpenAI-Compatible APIs Are Great Until Streaming Breaks: What I Check Before Switching Providers

Comments
7 min read
GLM 5.2 Is Now Available on TokenBay: Testing It with the OpenAI SDK

GLM 5.2 Is Now Available on TokenBay: Testing It with the OpenAI SDK

Comments
5 min read
One API Key for GPT, Claude, Gemini, and Qwen: A Practical Guide to OpenAI-Compatible Model Routing

One API Key for GPT, Claude, Gemini, and Qwen: A Practical Guide to OpenAI-Compatible Model Routing

Comments
5 min read
I Tested 6 AI API Gateways in 2026 — Here's My Real-World Comparison

I Tested 6 AI API Gateways in 2026 — Here's My Real-World Comparison

Comments
5 min read
How I Cut My AI API Bill by 40% Without Changing a Single Line of Application Code

How I Cut My AI API Bill by 40% Without Changing a Single Line of Application Code

Comments
4 min read
loading...