Google’s Gemini Omni 1.1 Flash Takes #1 in Text-to-Video, Plus 3 New Models Enter Top Rankings — AI Model Rankings for August 29, 2026

Daily AI Model Rankings Update — August 29, 2026

Google just shook up the video generation leaderboard: Gemini Omni 1.1 Flash has dethroned its predecessor to claim the #1 spot in text-to-video, while also crashing into the image-to-video top 10 — and a new code generation contender called hy4-preview quietly entered the rankings. If you’re building with video or code generation APIs, today’s changes are worth your attention.

What Changed Today

  • Text-to-Video — New #1: gemini-omni-1.1-flash overtook gemini-omni-flash to become the top-ranked text-to-video model (ELO: 1515)
  • Text-to-Video — New in Top 10: gemini-omni-1.1-flash entered the rankings for the first time (ELO: 1515)
  • Image-to-Video — New in Top 10: gemini-omni-1.1-flash entered the rankings (ELO: 1488)
  • Code Generation — New in Top 10: hy4-preview entered the rankings (ELO: 1633)

Text-to-Video: Gemini Omni 1.1 Flash Is the New King

This is a significant shakeup. Gemini Omni 1.1 Flash didn’t just enter the text-to-video leaderboard — it landed at #1 with an ELO of 1515, blowing past the previous leader Veo 3.1 Audio 1080p (ELO: 1392) by a massive 123-point margin. That’s not a marginal improvement; it’s a generational jump. The model also displaced its own predecessor, gemini-omni-flash, suggesting Google has been iterating aggressively on its omni-modal architecture.

For developers, the practical question is API access and pricing. Google’s Gemini Omni models are accessible through the Gemini API ecosystem — look for the model ID gemini-omni-1.1-flash. Pricing details haven’t been fully published yet, but given the “Flash” designation, expect this to be positioned as a fast, cost-efficient option rather than a premium tier. If you’ve been using veo-3.1 at $0.40/sec for text-to-video, this could be a meaningful cost reduction with better quality. Watch Google’s API docs closely for rate limits and availability.

FREE GUIDE

Stop Writing Design Specs by Hand

Get the free visual guide: how AI tools generate GAMP 5 documentation directly from your PLC and DCS exports. Used by Life Sciences engineers who are done doing it manually.

No spam. Unsubscribe anytime.

Image-to-Video: Gemini Omni 1.1 Flash Enters Here Too

Gemini Omni 1.1 Flash also broke into the image-to-video top 10 with an ELO of 1488, which would place it significantly above the current leader Grok Imagine Video 720p (ELO: 1402). This suggests the omni-modal approach — a single model handling text, image, and video in a unified architecture — is outperforming purpose-built video models across multiple tasks.

For builders who need both text-to-video and image-to-video capabilities, this is compelling: one model, one API integration, top-tier performance in both categories. The current image-to-video leader from xAI (grok-imagine-video-720p) still has strong community vote counts (13,668+), so it’s worth watching whether Gemini Omni 1.1 Flash maintains its lead as more votes come in. But early signals are very strong.

Code Generation: hy4-preview Enters at a Stunning ELO 1633

This one flew under the radar but might be the most consequential change today. hy4-preview entered the code generation top 10 with an ELO of 1633 — which, if the score holds up with more votes, would make it the new #1 code generation model, surpassing Claude Opus 4.6 (ELO: 1561) by 72 points. That’s a staggering margin in a category where the top models have been separated by single digits.

We don’t yet have confirmed pricing or provider details for hy4-preview, and the “preview” tag signals this is likely an early-access release. The current code generation leader, Claude Opus 4.6 from Anthropic, runs at $5/$25 per MTok via claude-opus-4-6. If you’re evaluating code generation models for production use, keep an eye on hy4-preview’s vote count — early ELO scores with limited votes can shift dramatically. But if this holds, we’re looking at a potential new top coding model that nobody saw coming.

Current Leaders at a Glance

Category #1 Model Provider ELO Score
Text-to-Video gemini-omni-1.1-flash 🆕 Google 1515
Image-to-Video grok-imagine-video-720p* xAI 1402
Code Generation claude-opus-4-6* Anthropic 1561

* Note: gemini-omni-1.1-flash (ELO: 1488) and hy4-preview (ELO: 1633) have entered the top 10 in image-to-video and code generation respectively with scores above the current listed leaders. Official #1 designations may shift as vote counts stabilize. We’ll update when confirmed.

So What?

Today’s changes point to two clear trends builders should act on. First, Google’s omni-modal strategy is paying off big. A single Gemini model just took #1 in text-to-video and entered the image-to-video top 10 — if you’re building any video generation feature, you should be testing gemini-omni-1.1-flash immediately, because consolidating your video pipeline into one model simplifies your stack and likely reduces costs. Second, the code generation leaderboard may be about to flip. The hy4-preview entry at ELO 1633 is eye-popping, but it’s a preview model with an unproven vote count — don’t rip out your Claude Opus integration yet, but absolutely add it to your eval suite. The actionable move today: spin up a quick benchmark of gemini-omni-1.1-flash for your video use cases, and put hy4-preview on your watchlist for code generation. The models that matter are shifting fast, and the builders who test early ship first.

Scroll to Top