Mimo 2.5 Didn’t Just Make Mistakes. It Changed How I Evaluate AI Models.

For the past few weeks, I’ve been building and experimenting with AI agents.

Not just chatbots, but agents that search the web, modify files, interact with Linux terminals, schedule cron jobs, orchestrate MCP tools, and automate real work.

Like many people, I initially evaluated models using the usual metrics:

  • Benchmark scores
  • Coding ability
  • Reasoning capability
  • Context window
  • Cost
  • Speed

Those metrics are useful.

But after spending enough time with autonomous agents, I’ve come to believe they don’t measure the thing I care about most.

Operational reliability. Continue reading Mimo 2.5 Didn’t Just Make Mistakes. It Changed How I Evaluate AI Models.

Why OpenCode Go Is Worth It — An Honest Take for Coders and Automators

I’ve been using OpenCode Go for about two weeks now, and it’s become my go-to for coding, AI agents, and automation. Here’s an honest take on whether it’s worth it.

What Is OpenCode Go?

A $5 first month, $10/month after subscription that gives you access to 13 top open coding models. No per-token charges — just a flat fee with usage limits.

Subscribe to OpenCode Go

The Models

  • DeepSeek — V4 Pro, V4 Flash
  • GLM — 5.2, 5.1
  • Kimi — K2.7 Code, K2.6
  • Qwen — 3.7 Max, 3.7 Plus, 3.6 Plus
  • MiniMax — M3, M2.7
  • MiMo — V2.5, V2.5 Pro (with vision!)
Continue reading Why OpenCode Go Is Worth It — An Honest Take for Coders and Automators