AI benchmarks and playable demos

AI model comparisons, playable demos, and selected projects.

Concrete examples of how different AI models handle the same assignments, with original outputs, practical evaluations, and demos you can try directly.

Model benchmarks

The same assignment, different results.

Each benchmark preserves the original model outputs, screenshots, practical evaluation, and playable demos.

Blog

Notes from real-world work

Longer write-ups about software development, local models, and experiments tested on real projects.

A busy intersection with pedestrians, traffic, and a motorcycle in the GTA game by Claude Fable 5.1
  • Article

Fable 5.1 blew my mind

The first prompt already produced a huge, lively city. Another 52 corrective prompts over one weekend refined it into a fun, playable game.

  • September 2026
  • Claude Fable 5.1
  • Browser game
  • GTA benchmark

Read the article

A laptop running an AI coding agent and connected to a Mac M3 Ultra used as a local LLM server
  • Article

DeepSeek V4 Flash 0731 DS4 in real software development

Two practical tests of DeepSeek, Qwen, and Laguna on real tasks from an established codebase. The same baseline and prompt within each test, with results verified in the application.

  • August 2026
  • Local LLMs
  • C#
  • Mac M3 Ultra

Read the article

Projects

Other projects

  • Live

flatscraper.com

A new-build property search engine for Slovakia. It tracks more than 600 active residential development projects and updates them through automated data collection.

  • Web
  • Data
  • Real estate

Open flatscraper.com

  • Planned

UIMD

A toolkit for building terminal UI applications. Python and C++ versions for macOS are in progress; the goal is to make text interfaces for developer tools easier to build.

  • Terminal UI
  • SDK
  • C++
  • Python

View on GitHub

Author and contact

Marek Dubovsky

A software developer with experience across desktop, mobile, and web, mainly in C/C++, Python, and JavaScript.

FAQ

Frequently asked questions

What is Dubovsky Labs?

Marek Dubovsky's personal site for practical experiments, AI model benchmarks, playable demos, and selected software projects.

Can benchmark results be tried directly?

Yes. Playable and interactive results include direct links to demos and 3D viewers.

Are original model outputs modified?

Original benchmark artifacts remain preserved. Any fixes, follow-up prompts, or website-owned wrappers are disclosed with the result.

What projects are included?

Alongside AI benchmarks, the site presents selected desktop, mobile, and web work, including UIMD and flatscraper.com.