FOTOhub launches IDA Q 1.0, our first proprietary image generation model

Mateusz Ulewicz, Founder & CTO · · 9 min read

FOTOhub launches IDA Q 1.0, our first proprietary image generation model, the third pillar of the platform alongside Gabriel AI and IDA Voice & Audio. Top 5 worldwide on the DesignArena benchmark, 20x cheaper than GPT-Image-2, free for 7 days.

FOTOhub launches IDA Q 1.0, our first proprietary image generation model

FOTOhub launches IDA Q 1.0, our first proprietary image generation model

Category: Product Launch / AI / Technology

Date: July 24, 2026

Author: Mateusz Ulewicz, Founder & CTO

Tags: IDA Q 1.0, proprietary AI model, image generation, product launch, Gabriel AI, IDA Voice & Audio, AWS

Summary

FOTOhub.app is launching IDA Q 1.0 today, our first fully proprietary text-to-image generation model, built and hosted on FOTOhub infrastructure in partnership with AWS. It becomes the third technological core of the platform, alongside Gabriel AI (the orchestrator that understands user intent) and IDA Voice & Audio (our own voice and music engine). The model ranks in the top 5 worldwide on the DesignArena benchmark, delivers precise text rendering on images, native multilingual support, and pricing 20x lower than competing GPT-Image-2. For the next 7 days after launch, every FOTOhub user, including those on the free plan, generates images with IDA Q 1.0 at no cost.

Why this is the day we've been waiting for

Two years ago, when I started building FOTOhub, we had one strategy: integrate as many of the world's best AI models as possible into one place, so users wouldn't have to choose between dozens of platforms. That strategy worked, today we have over 200 models from 10+ providers and 750,000 users in 40+ countries.

But an integrator is not the same thing as a technology creator. Somewhere along the way we asked ourselves a question that now defines FOTOhub's next chapter: what happens if we start building our own models, tuned precisely to how our users actually create?

The answer is the IDA model family, our own line of AI technology, hosted on our own infrastructure, developed by our own team. IDA Voice & Audio already powers voice generation, voice cloning, and music for thousands of users. Today it's joined by IDA Q 1.0, an image generation model that puts FOTOhub in the group of companies that don't just integrate AI, they build it.

"We don't want to just be the best place to use other companies' AI. We want to be a company that builds its own technology, one that others will build on." Mateusz Ulewicz, Founder & CTO

Three pillars of the platform

With IDA Q 1.0 launching, FOTOhub's architecture now rests on three proprietary pillars that together add up to more than the sum of their parts:

1. Gabriel AI, orchestrator

Understands the user's natural language, decides which model and module best matches the intent, and executes the task end to end. The brain of the system.

2. IDA Voice & Audio, voice and music

Our own text-to-speech engine, music generation, and voice cloning. 50+ voices, 30+ languages, a full dubbing pipeline. Zero dependency on external audio providers.

3. IDA Q 1.0, image generation

The new third pillar. A text-to-image model designed from the ground up around precision: accurate text placement, composition control, and multilingual support within a single generation.

These three pieces don't operate in isolation. Gabriel AI already knows when to route a user's request to IDA Q 1.0, when to send it to an external model, and when to combine multiple engines into a single task. That's the difference between "we have access to 200 models" and "we have a system that thinks for you about which model to use."

What IDA Q 1.0 actually does

IDA Q 1.0 is a text-to-image model optimized around the three things that most often break the output of other generators: text on images, composition, and prompt fidelity.

Precise text rendering

Most image generation models still struggle with clean, correct text: typos, distorted letters, misplaced headlines. IDA Q 1.0 treats text as a first-class compositional element, not an afterthought. Poster headlines, product labels, quotes inside graphics come out clean and legible, exactly as shown in our "FUTURE OF CREATION" poster example in the gallery below.

Composition control through structure

The model doesn't "guess" where things belong on the canvas. Instead, it decomposes a scene into a background and separate elements, each with its own description, position, and relationship to the others, the same way a designer actually thinks about composing a poster or a photograph. The result: fewer random artifacts, more control over the final frame.

Native multilingual support

The prompt language and the text rendered on the image can differ. The model correctly renders Polish diacritical marks (ą, ę, ń, ś, ź, ż) just as well as the standard Latin alphabet. In our coffee cup example, the handwritten note "Dzień dobry" came out flawless, letter by letter.

FOTOhub's prompt engine, our own intelligence layer

This is the part we're especially proud of, and one our team built specifically for IDA Q 1.0. Before a user's prompt reaches the model, it passes through our proprietary engine built on Claude Haiku via AWS Bedrock, which:

This isn't a cosmetic feature. It's the difference between an average result and a result that actually reflects what you had in mind, without needing to learn prompt engineering.

Benchmarks: where we actually stand

We don't like inflated PR numbers, so here are the figures as they are. On the independent DesignArena benchmark (Elo rating, measuring generation quality against real-world design tasks), IDA Q 1.0's underlying engine ranks 5th in the world with a score of 1285 Elo, just below the market leaders, and clearly ahead of a long list of established models:

ModelElo (DesignArena)
GPT Image 21405
GPT-Image-1.51327
Gemini 3.1 Flash Image Gen 2K1318
Gemini 3.1 Flash Image Gen1310
IDA Q 1.0 (underlying engine)1285
Gemini 3 Pro Image Gen 2K (Nano)1284
Gemini 3 Pro Image Preview1259
UNI-1.11254
Recraft V4.1 Utility Pro1245
Krea 2 Medium1245
FLUX.2 [flex]1244
FLUX.2 [pro]1239
Seedream Lite 5.01236
Krea 2 Large1235
Imagen 4 Ultra Generate Preview1233

What does this mean in practice? IDA Q 1.0 clearly beats Recraft, both Krea 2 tiers, FLUX.2, Seedream Lite, and Imagen 4 Ultra, models that are today's standard across many design tools. It's not yet at the level of Gemini Flash or GPT-Image-2, and we say that openly, without spinning the numbers. But for a model just entering the market under our own brand, 5th place worldwide on a respected benchmark is a very strong start. We'll keep improving it, and this is only the first version.

Pricing and availability

ParameterValue
Price after the free period0.5 credit / image
GPT-Image-2 price (comparison)10 credits / image
Difference20x cheaper
Resolutions1K (1024×1024), 1.5K (1536×1536), 2K (2048×2048)
Generation time~30s (1K) / ~90s (1.5K) / ~3.5 min (2K)
Free period7 days from launch (through July 31, 2026)
AvailabilityAll plans, including free

The 20x lower price isn't a coincidence, it's a direct result of the model running on our own GPU infrastructure, with no per-request licensing fees paid to a third-party API. When you build your own technology, you stop paying someone else's margin.

Gallery: what IDA Q 1.0 can really do

The images below were generated directly by IDA Q 1.0, with no manual curation or repeated attempts. These are the first results for each prompt, passed through the prompt engine described above.

1. Portrait in natural light

Prompt: "Kobieta z bukietem polnych kwiatów, stojąca na tle złotego krajobrazu o zachodzie słońca, elegancka sukienka, delikatny wietrzyk we włosach, fotorealistyczne, ciepłe światło godziny magicznej" (Polish: "a woman with a bouquet of wildflowers, standing against a golden landscape at sunset, elegant dress, gentle breeze in her hair, photorealistic, warm golden-hour light")

!Woman with a bouquet of wildflowers in golden sunset light

Natural golden-hour light, vivid bouquet colors, realistic background depth of field, with zero manual retouching.

2. Night scene with neon

Prompt: "Nowoczesne sportowe Porsche sfotografowane w ciemności, dramatyczne neonowe oświetlenie w odcieniach fioletu i cyjanu, mokra jezdnia odbijająca światła, miejska sceneria nocna, futurystyczny styl 2026, ultra szczegółowe" (Polish: "a modern sports Porsche photographed in darkness, dramatic neon lighting in purple and cyan tones, wet road reflecting the lights, urban night setting, futuristic 2026 style, ultra detailed")

!Sports Porsche on a wet street under neon night lighting

Reflections on wet asphalt, atmospheric depth, precise bodywork detail, exactly the level expected from a product campaign.

3. Text rendering: poster

Prompt: "Minimalist modern poster design with bold typography, headline text FUTURE OF CREATION in large clean sans-serif letters, geometric abstract shapes in deep blue and orange gradient background, professional graphic design, high contrast"

!Minimalist poster with the headline FUTURE OF CREATION

Zero typos, zero distorted letters. This is exactly the area IDA Q 1.0 was built to win.

4. Architectural interior

Prompt: "Luksusowe minimalistyczne wnętrze salonu z dużymi oknami, naturalne światło dzienne, drewniane akcenty, betonowe ściany, designerskie meble skandynawskie, wysokie sufity, architektura nowoczesna, fotorealistyczne wnętrze" (Polish: "a luxurious minimalist living room interior with large windows, natural daylight, wooden accents, concrete walls, Scandinavian designer furniture, high ceilings, modern architecture, photorealistic interior")

!Minimalist living room with concrete walls and large windows

Realistic perspective, correctly proportioned furniture, natural light falling through the windows, ready to use in a real estate listing or an architect's portfolio.

5. Fantasy world

Prompt: "Fantastyczne miasto pływające wśród obłoków, kryształowe wieże połyskujące w świetle zachodzącego słońca, latające statki między budynkami, styl epicki fantasy, bogata kolorystyka, malarska iluzja głębi" (Polish: "a fantastical city floating among the clouds, crystal towers glinting in the light of the setting sun, flying ships between the buildings, epic fantasy style, rich color palette, painterly illusion of depth")

!Fantastical crystal city floating among the clouds

Beyond photorealism, the model handles pure creative fiction equally well, depth and light do the heavy lifting here.

6. Polish text on an image

Prompt: "Filiżanka gorącej kawy na drewnianym stole w kawiarni, obok kartka papieru z odręcznym napisem Dzień dobry, poranne światło wpadające przez okno, ciepła atmosfera, fotorealistyczne" (Polish: "a cup of hot coffee on a wooden café table, next to a piece of paper with a handwritten note saying good morning, morning light streaming through the window, warm atmosphere, photorealistic")

!Cup of coffee with a note reading Dzień dobry (good morning)

Polish diacritical marks, the "ń" in "Dzień", rendered correctly, with no distortion. For a platform with Polish roots, that's a detail we particularly cared about getting right.

Our partnership with AWS reaches a new stage

Amazon Web Services has been our investor and strategic partner since Q2 2026, first as part of a funding round, then as an AWS Portfolio Company with access to architectural support and the latest cloud technologies. Until now, that partnership mainly meant integrating AWS models (Bedrock, Nova) into our catalog and running infrastructure for our 12 modules.

IDA Q 1.0 is the natural next step in that relationship: FOTOhub's first proprietary model, running on AWS GPU infrastructure from day one, with its prompt engine powered by Claude via AWS Bedrock. It shows where our AWS partnership is heading, from "we integrate your models" to "we build our own AI technology on your infrastructure, from the first line of code." For a company like FOTOhub, building proprietary models without confidence in infrastructure scalability would be risky. Our partnership with AWS removes that risk.

What's next

IDA Q 1.0 is a first version, not a final one. In the coming months we're working on higher maximum resolution, expanded support for image editing (not just generation from scratch), and deeper integration with Gabriel AI, so the orchestrator can automatically decide when IDA Q 1.0 is the right choice for a given task, and when one of our 200+ integrated models is a better fit.

The IDA family, Voice & Audio, and now Q 1.0, is the foundation we'll keep building on. This is a strategy, not a one-off project: FOTOhub is moving from a company that integrates other people's AI exceptionally well to a company that builds its own technology.

IDA Q 1.0 is available today for all FOTOhub users, including the free plan, and free for the next 7 days. Log in at fotohub.app, select IDA Q 1.0 in the image generator, and see for yourself.

We're building our own technology. And sharing it with everyone who creates.

Mateusz Ulewicz

Founder & CTO, FOTOhub

Top 5 / 2.1M in the main Crunchbase category

FOTOhub, Creative AI OS

fotohub.app | [email protected]

Crunchbase: crunchbase.com/organization/fotohub

Investor Relations: fotohub.app/investors

FOTOhub, 750K+ Users | $8.6M Raised | $180M Valuation | AWS Portfolio Company