DEV Community

Riley Li profile picture

Riley Li

Data engineer building scalable pipelines and tinkering with AI-assisted workflows.

Location London, UK Joined Joined on 
Inventory What You Cannot Pin Before You Choose Free Agent Compute

Inventory What You Cannot Pin Before You Choose Free Agent Compute

Comments
8 min read
Budget the Agent Loop Before You Choose Shared or Isolated Compute

Budget the Agent Loop Before You Choose Shared or Isolated Compute

Comments
7 min read
Score the Host Before You Promote a Vibe-Coded Agent

Score the Host Before You Promote a Vibe-Coded Agent

Comments
6 min read
Choose Agent Compute by Replayable Evals, Not by Spare Capacity

Choose Agent Compute by Replayable Evals, Not by Spare Capacity

Comments
8 min read
Pass Four Fit Gates Before You Choose Free or Isolated Agent Compute

Pass Four Fit Gates Before You Choose Free or Isolated Agent Compute

Comments
7 min read
Who Owns the Failed Agent Run? Score Free Shared Compute Against a Box You Control

Who Owns the Failed Agent Run? Score Free Shared Compute Against a Box You Control

Comments
7 min read
Free Compute or Self-Hosted? A Break-Even Guide for Agent Workloads

Free Compute or Self-Hosted? A Break-Even Guide for Agent Workloads

Comments
5 min read
Rank Agent Runtimes by the Contract You Can Enforce

Rank Agent Runtimes by the Contract You Can Enforce

Comments
8 min read
Invert Non-Negotiable Constraints Before You Accept Free Agent Compute

Invert Non-Negotiable Constraints Before You Accept Free Agent Compute

Comments
7 min read
Keep the Agent Loop Portable, Then Choose Free Shared Compute

Keep the Agent Loop Portable, Then Choose Free Shared Compute

Comments
7 min read
Pick Agent Compute by Blast Radius, Not by the Price Tag

Pick Agent Compute by Blast Radius, Not by the Price Tag

Comments 1
7 min read
Choose Agent Compute by What You Can Falsify

Choose Agent Compute by What You Can Falsify

Comments
7 min read
Score the Agent Before You Pick the Endpoint

Score the Agent Before You Pick the Endpoint

Comments
10 min read
Score the Lane First: Free Shared Workspace, Paid API, or Self-Hosted

Score the Lane First: Free Shared Workspace, Paid API, or Self-Hosted

Comments
7 min read
Designing for a 10M-Token Ceiling: A Token-Budget Scheduler for Background LLM Jobs

Designing for a 10M-Token Ceiling: A Token-Budget Scheduler for Background LLM Jobs

Comments
5 min read
A Token Bucket for Free-Tier LLM Endpoints: Rate Limiting Without a Fancy Gateway

A Token Bucket for Free-Tier LLM Endpoints: Rate Limiting Without a Fancy Gateway

Comments
4 min read
Keep the Regex Writer Until Shadow Receipts Match

Keep the Regex Writer Until Shadow Receipts Match

Comments
10 min read
Don't Let a Free LLM Tier Vanish: A Self-Healing Token Budget in 80 Lines

Don't Let a Free LLM Tier Vanish: A Self-Healing Token Budget in 80 Lines

Comments
5 min read
Free Models and a Free Server: A 100-Line LLM Failover Proxy

Free Models and a Free Server: A 100-Line LLM Failover Proxy

Comments
5 min read
Free Tokens or Your Own GPU? A Practical Fit-Test for AI Workloads

Free Tokens or Your Own GPU? A Practical Fit-Test for AI Workloads

Comments
5 min read
Free LLM Tiers vs. Self-Hosted Inference: A Four-Score Script That Picks for You

Free LLM Tiers vs. Self-Hosted Inference: A Four-Score Script That Picks for You

Comments
4 min read
Free Models and a Free Server: A 30-Line Budget That Lasts the Month

Free Models and a Free Server: A 30-Line Budget That Lasts the Month

Comments
4 min read
A 60-Line Probe That Decided Between a Free Server and My Own Hardware

A 60-Line Probe That Decided Between a Free Server and My Own Hardware

Comments
5 min read
Free LLM Tier or Self-Hosted? Three Numbers Decide, Not the Price Tag

Free LLM Tier or Self-Hosted? Three Numbers Decide, Not the Price Tag

Comments
5 min read
Token Meters for Free LLM Tiers: Three Leaks My 150-Line Logger Caught in an Hour

Token Meters for Free LLM Tiers: Three Leaks My 150-Line Logger Caught in an Hour

Comments
7 min read
Debugging a Flaky LLM Pipeline: Timeouts, Truncation, and a 40-Line Probe

Debugging a Flaky LLM Pipeline: Timeouts, Truncation, and a 40-Line Probe

Comments
5 min read
Three Nights of Empty LLM Responses — and the Isolation Trick That Finally Caught the Culprit

Three Nights of Empty LLM Responses — and the Isolation Trick That Finally Caught the Culprit

Comments
4 min read
Minimax H3 Is Having a Moment. My Prompt Set Doesn't Care.

Minimax H3 Is Having a Moment. My Prompt Set Doesn't Care.

Comments
4 min read
A Credit-Card-Free Triage Bot for Noisy CI Logs

A Credit-Card-Free Triage Bot for Noisy CI Logs

Comments
5 min read
Every Week a New Model Is "Cheaper and Better" — Here's the 30-Minute Harness That Settles It

Every Week a New Model Is "Cheaper and Better" — Here's the 30-Minute Harness That Settles It

Comments
5 min read
I Stopped Reading Model Release Threads and Built a Release-Day Eval Ritual Instead

I Stopped Reading Model Release Threads and Built a Release-Day Eval Ritual Instead

Comments
5 min read
loading...