BuildnWriteBlogProductAbout BuildnWrite ›

Categories

  • All posts 100
  • Guides 28
  • Concepts 14
  • Insights 54
  • Experiments 4

Topics

  • n8n automation 14
  • Integrations and auth 2
  • Deploy and servers 1
  • Data and APIs 1

#model-comparison

3 posts tagged model-comparison

  • Experiments

    Which AI Model for Work Automation? GPT-6 Luna vs Claude Haiku 5.5 on 11 Real Tasks

    11 real work tasks, from diagrams and slides to browser work, ran through five GPT and Claude models, three runs each, with scores and costs.

    October 8, 20268 min read

  • Experiments

    Aside Browser: Should You Connect Claude or GPT? Accuracy, Speed, and Cost Test

    I ran three web tasks nine times each on Aside with Claude and GPT models. Three models scored perfectly; the cheapest, Luna, missed twice on shopping.

    October 6, 20268 min read

  • Insights

    Fable 5 vs Opus 4.8: Real Session Cost and Agent Behavior Compared

    Fable 5 costs twice as much per token as Opus 4.8, but measured sessions cost 1.0 to 1.4 times as much. See the token, cost, and intervention data behind it.

    June 10, 20266 min read

← All posts

Hands-on notes on AI agents and work automation.

BuildnWrite · Seoul, South Korea · About us · RSS

© 2026 BuildnWrite. All rights reserved.