Mashdun
AI GrowthWorkCapabilitiesIntegrationsProcessLabsMarketplaceBlogAbout
Get in Touch
Mashdun

Full-stack web developer, AI engineer & growth marketer. Building production-grade apps and intelligent solutions.

Navigation

  • Portfolio
  • Capabilities
  • Process
  • Labs
  • About

Resources

  • Blog / Notes
  • Marketplace
  • Contact
  • RSS Feed

Legal

  • Terms of Service
  • Privacy Policy
  • Cookie Policy
  • Do Not Sell or Share
  • Delete my data

© 2026 Mashdun. All rights reserved.

Built with Next.js, Tailwind CSS & Prisma

    All Labs
    AI
    Multi-model
    LLM-as-judge
    OpenAI
    Claude

    AI Article Arena

    Multi-model generation with an LLM-as-judge scoring loop: choose two or more models (OpenAI + Claude), each writes a short article on your topic in parallel, then an independent judge model grades every article — blind to who wrote it — on clarity, accuracy, engagement, and structure, and crowns a winner.

    AI Article Arena

    Multi-model
    LLM-as-judge
    OpenAI
    Claude

    No models added yet — add at least two to run the arena.

    Judge = Claude Opus 4.8 when configured, otherwise GPT-4o.

    How it works

    1. Pick a provider, choose from its available models, and add two or more entrants
    2. Each model writes a short (~180-word) article in parallel
    3. A separate judge model scores every article (blind to which model wrote it)
    4. Scores rank clarity, accuracy, engagement & structure — and crown a winner

    From real client work

    This is the pattern behind reliable AI writing pipelines: don't trust one model's output — generate with several, then let an independent judge model grade them against an explicit rubric. The same LLM-as-judge loop drives the dashboard's article builder and our eval harnesses.

    See the full benchmark arena