back to top
HomeTechGoogle's Gemini Omni Can Write Math on a Chalkboard. AI Video's Hardest...

Google’s Gemini Omni Can Write Math on a Chalkboard. AI Video’s Hardest Problem May Be Getting Easier

- Advertisement -

Google hasn’t announced Gemini Omni. A reddit user just found it anyway.

Someone opened their Gemini app, got a pop-up for a model they’d never heard of, and started generating video. What came out has been making rounds on Reddit for the past few days, mostly because of one clip.

The chalkboard video is why.

A professor writing a full mathematical proof in chalk, narrating as he goes, text legible, delivery natural, physics mostly convincing. AI video has never handled text well. This one does. And it’s not just text, the audio, the movement, the realism are all hitting at a level that has people genuinely uncomfortable in a way the usual AI video demos don’t.

It’s a leak. Google hasn’t said a word. I/O is next week. But whatever Omni is, the early results suggest something shifted.

What we know about Omni

Not a lot, officially. The user got a pop-up in the Gemini app prompting them to “create with Gemini Omni”, described as a new video generation model with the ability to remix videos, edit directly in chat, and use templates. That’s the entire official description. Google hasn’t confirmed anything exists.

Max Weinbach dug into the metadata and found that Omni appears to be an extension of Veo rather than something built from scratch. Which tracks. Google has been developing Veo for a while and the output here looks like a significant step forward from what Veo was producing, not a completely different direction.

I/O is next week. That’s almost certainly when this becomes official and we get actual details on what Omni is and how it fits into the broader Gemini lineup.

Where it still lacks

The chalkboard result is impressive. The spaghetti test is a different one.

The original Will Smith prompt got blocked by Omni’s guardrails, so the user rewrote it, two men at a seaside restaurant, white tablecloth, approaching the table and eating spaghetti while having a conversation.

Spaghetti appears from nowhere on plates that were empty seconds earlier. The eating doesn’t match the bites. The inconsistencies that the chalkboard video mostly avoids stack up quickly here. Another Reddit user ran the same prompt through ByteDance’s Seedance 2 and got a noticeably more consistent result.

So Omni isn’t uniformly ahead of everything. The text handling is genuinely new. Physical realism on complex interactions still has the rough edges you’d find elsewhere.

You May Like: daVinci-MagiHuman Finally Makes Open-Source AI Video Feel Real

The usage question most ignore

Those two generations, the chalkboard and the spaghetti consumed 86% of this user’s daily quota on a Google AI Pro plan. There was some Gemini Flash usage the same day so it’s not a perfectly clean number, but the direction is clear. Video generation on Omni is expensive in terms of quota, and that’s going to be the conversation nobody is having right now but everyone will be having the moment Google makes this official and people hit their limit inside the first two prompts.

What comes next

Google said video is here to stay after OpenAI shut down Sora earlier this year. Omni looks like the proof of that commitment. The chalkboard result alone suggests something real has shifted on the text rendering problem, even if the model isn’t consistent across all prompts yet.

We’ll know the full picture next week. Until then the chalkboard video is the thing worth watching twice.

Don’t miss any Tech Story

Subscribe To Firethering NewsLetter

You Can Unsubscribe Anytime! Read more in our privacy policy

LEAVE A REPLY

Please enter your comment!
Please enter your name here

YOU MAY ALSO LIKE
Claude Will Soon Leave a Hidden Mark on Everything It Writes

Claude Will Soon Leave a Hidden Mark on Everything It Writes

0
Anthropic is adding invisible watermarks to text generated by Claude, and unlike a visible label, the marking is designed to travel with the text when users copy and paste it elsewhere. The move comes as AI-generated content becomes harder to distinguish from human writing and as the European Union begins requiring AI companies to make generated or manipulated content machine identifiable. Anthropic says the marking happens at the model level, meaning it can follow Claude-generated text across different products and surfaces rather than being tied to a particular app. The company also says it may survive some editing. But it raises a question, can AI-generated text actually be made traceable once it leaves the model that created it? And Anthropic's approach suggests the answer may be more complicated than simply adding a hidden signature to every sentence.
Zuckerberg Wrote 14 Pages About Open AI. His Best AI Model Is Still Closed

Zuckerberg Wrote 14 Pages About Open AI. His Best AI Model Is Still Closed.

0
Mark Zuckerberg published a 14-page essay today about why open-source AI is the path forward for humanity. Distribute intelligence rather than centralize it. Put the power in everyone's hands. A new era of personal empowerment. On the same day, Meta released Muse Glimmer, an open-source version of its most powerful model, Muse Spark, that anyone can download, modify, and build on for free. But the interesting part is, Muse Spark itself stays closed. You still pay to access it. The open version is nearly identical, Meta says, but the model that actually competes at the frontier, the one Zuckerberg's essay is implicitly defending remains behind a paywall. That gap between the philosophy and the product decision is what makes today's announcement interesting.

The Biggest AI Companies Are All Building Their Own Chips. That’s Not a Coincidence.

0
Anthropic confirmed this week it's hiring a custom silicon team to design chips for running Claude. The announcement was quiet a job listing, a spokesperson confirmation, no big launch event. Easy to file under "interesting but expected" and move on. But zoom out for a second. OpenAI shipped its first custom inference chip in June. Google has been running models on its own TPUs for years. Meta has designed and deployed its own silicon. Mistral is reportedly exploring the same path. And now Anthropic. Five of the most important AI labs in the world, all arriving at the same decision, within roughly the same window. None of them are copying each other. All of them looked at the same competitive landscape and reached the same conclusion independently. That kind of convergence doesn't happen by accident. It happens when an entire industry agrees that the thing everyone assumed was someone else's problem is actually the problem and that whoever solves it first has an advantage that's very hard to close later.