I heard a radio spot recently and I wondered if the voice was a real person or AI. It makes we wonder how such industries are dealing with this gen-AI revolution. We spend a lot of time here thinking about how it affects software developers, but I hardly ever see any commentary on how it is affecting screen and voice actors.
Everyone will probably go through the 5 stages of grief w.r.t AI adoption in their field and the gatekeepers will (rightfully) hold onto hallucination and errors as reasons to delay incorporation or to incorporate it with more human handholding.
Software developers, unfortunately, have been convinced that they don't need to unionize, so have no collective bargaining power for dealing with situations like this.
In my experience this is what almost everyone thinks, right up to the point where it starts happening to them.
And it isn't just an individual blind spot, organizations suffer from the same thing in the sense that many companies will be happy to automate away all their labor to avoid paying the "human tax" without giving much thought to the fact that the need for the company itself will also be automated away soon after.
If you can replace nearly all of your workers with some cheap tokens and prompts, everyone else who previously would have been a customer can replace your whole company with the same thing.
> it's hilarious to see some people here go on about how they're proud to work so that they can be replaced
To be fair, most of us spent our entire career trying our best to automate ourselves away one way or another, and always seen that as our job description.
No, that is just nerd heroism lore that indeed has always shown up in comments, that part is definitely true.
Some people dislike automation altogether and like the computational aspect. Most people seem to enjoy constructing virtual worlds and Rube Goldberg machines.
I thought HN was mostly people who hated doing 'the thing' and would procrastinate until they build a system that does 'the thing' and then they would work tirelessly to never do 'the thing' themselves ever again.
I think I sit somewhere between these two descriptions. I've always supported unions and other "for the common good" type machinery. At the same time, I desperately also don't want to end up doing something the equivalent of barring the use of calculators just so I can toil away at a 9-5 crunching numbers more slowly instead. If AI really does replace all of the meaningful jobs we can do... great - I'd rather we make sure the spoils of that production are distributed than try to cling to preventing the technology from being used.
Inexplicably for the current moment, AI has so far actually meant I have even more to do instead of less. I expect that to change eventually... but man, is it a bit of whiplash to go from wondering if my job will be around in 10 years to starting the work day and being more backed up than ever. Doubly so since the rate of change does not seem to be very evenly distributed by tech role.
Agreed, but I must've severely messed up my message if that was supposed to be a spoiler. To be clearer: I think the unions, economic policy, and such should help with that part rather than us wield the very same to hide from replacing the work we do with something much easier. I.e., don't be ashamed to work towards that, just come together to ensure we all share the spoils.
What options do you have? If you're the 10x engineer you'll be a 100x one and do just fine. Otherwise what can you do besides holding on as long as you can or start searching for a new career.
I agree with your point, but I don't think it's funny at all. The whole field is changing because of this. The career, except for those ones lucky enough to have jobs where they're valued and they can decide how to use AI tools, is turning into shit because of this and software engineers have become a lot less valuable as management is trying to turn them into reverse centaurs.
Half of them are from India or China. Racial diversity is known to be detrimental to unionization. In fact, Amazon used this exact strategy to bust nascent Somali unions.
Now, they cannot prevent offshoring or AI use or replacements. Very very rarely if ever.
But they can negotiate a large layoff and give better pay/conditions during work etc. This view that unions can prevent anything... when so many jobs have gone to east EU (all amazing and good at what they do but lets face it, it is due to pay) so yea there wont be any union preventing AI...
Personally, I just can’t find it in me to be that self-interested. It’s how other people oppose buildings near them because it blocks their view. I really don’t want to stop other people from writing software. If they want to use AI to do it so be it.
That’ll be somewhat detrimental to me perhaps (which isn’t certain), but that’s okay. I’d rather attempt to adapt to a changing world.
Thats great, although understand that unionizing protects the more vulnerable of also the software engineered. You not advocating for your rights also means weakening others. Is it still self interested to unionized from that perspective?
You may be surprised to know that residential areas in San Francisco are height limited, sometimes as short as 4 stories. They're not exactly anti-progress.
Ok SF needs somewhat higher-density housing. But wanting something good for yourself and your peers isn't anti-progress. If the benefits accrue to only a few people, it's not progress.
I don't know that unions would necessarily negotiate for no use of AI. The SAG-AFRA deals don't preclude all use of AI; they just requre consent and negotiation in certain cases.
A union doesn't give you unilateral power; it just gives you a better seat at the bargaining table. Capital pools its resources to negotiate better as a single bloc; why shouldn't labor as well?
> Personally, I just can’t find it in me to be that self-interested.
Yes, exactly. Do not advocate for yourself and your colleagues by joining them in a united front. Your voice does not matter. You will continue to adapt to widening wage disparities.
You will accept a pay cut and increased productivity goals and be happy to continue adapting as your cost of living keeps on climbing.
This is the first time I've heard someone argue that unionizing is a self-interested pursuit. It's called "collective bargaining" for a reason; workers who cooperate in negotiations with their employers have more leverage than those who negotiate individually. The entire point is to work together for the betterment of the group!
as an art enjoyer, i do not feel that having less media written, drawn, or voiced by a computer is a loss.
a computer can never be sad or horny or have taste. it can only make product. i also think artists should be paid living wages with good working conditions.
thank you SAG-AFTRA, WGA, IATSE, AEMI and Teamsters.
See current situation where the Carpenters' Union in California is pushing hard against an initiative that would make it much easier to build more housing in the state because it would also reduce requirements for union labor.
Unions are very much an in-group vs out-group phenomenon (and in many cases, the benefits are specifically to the more senior union members vis a vis the less senior ones.)
> See current situation where the Carpenters' Union in California is pushing hard against an initiative that would make it much easier to build more housing in the state because it would also reduce requirements for union labor
If you're talking about AB 1751 (Missing Middle Townhome Ownership Act) I think you have this exactly backwards.
> Unions are very much an in-group vs out-group phenomenon
As opposed to what? You and the GGP phrase this as a dichotomy, but I'm really curious what the other side of this is, because "not being in a union" doesn't erase your self interest or all the many many groups of people doing the same that don't feel like they're apparently being overly self interested by forming corporations, governments, non-profits, advocacy groups, etc, etc, etc
Have you ever not been able to send your kids to school for weeks (or years in the case of COVID) because the teachers refused to work? Have you ever been unable to get timely medical treatment because nurses refused to work? Have you ever been unable to work yourself because the transit workers refused to work?
People have this right to take better jobs at other companies. This is how we get better working conditions: competition between employers. Tech has figured this out. I’ve never felt the need to go on strike.
Unions can be a deterrent to this. If you have sectoral bargaining, you can’t get a better deal anywhere because everyone pays the same. Or even if it’s within a single company, you’re just stuck with whatever scheme the union bosses agreed to.
But also, unions are a deterrent to new employers since fewer people want to start companies in industries where unions have taken over. They’re a nightmare to deal with. Fewer employers means less competition means worse services.
We have a dynamic economy with lots of competition. It’s easy to switch jobs and find a better place to work. We should not want to give this up.
You realize his examples are public sector unions. Which mean the “employer” you are referring to is the citizens? Not some evil corporation, government mandated monopolies which then get taken over to serve their union members.
I don’t have any objection to voluntary unions. Meaning people can join them for support or collective bargaining, but that citizens can just opt to not employ anyone in the union. Unions could be support structures for workers, instead they are mafia bosses.
Well… yeah. But I didn’t blame the employees, I blamed the employer who must have seen the writing on the wall for a long time and yet let it come to this.
I can't speak for the OP, but I don't think their point was that they're against the self-interested nature of unionizing.
I read it more like they're against trying to stop, hinder, or prevent the use of new technology, simply because it's personally threatening.
And I feel the same way. If I was part of a union, I would believe in the collective bargaining power, but I wouldn't necessarily believe we should use that to pursue every single possible agenda that might benefit us. And I've personally pooled my capital with others in the past in order to achieve certain outcomes. And in those situations I also didn't attempt to reach so far as to limit the totally valid freedoms of others just because I knew it might be better for me.
Acting in the interest of a collective doesn't mean one must also abandon the ethics they usually employ when acting in their solo self interest.
>Personally, I just can’t find it in me to be that self-interested
Caring about your ability to pay your bills, and support your family is "self-interested" in the same way that choosing to not drive into oncoming traffic is "self-interested".
Has it occurred to you that it isn't only self-interest but concern for others that drives collective bargaining? It's essentialy a way for the little guys to stand up to the 1% to fight against the unending series of abuses and diminishment of the average person's ability to earn a livable wage and work under fair conditions.
Yes. But there's a clear difference between, say, stealing someone's voice, identity or intellectual property, and preventing someone from using automation to do labor.
The former is obviously unfair to the person whose identity is stolen. The latter is sour grapes from people who want to control the future.
I also can't bring myself to get incensed over someone using a robot to write code. It's literally the modern equivalent of smashing knitting looms.
Yeah, the problem is, someone has to start the union. It takes work. You have to get the whole workforce to vote on it. A lot of software engineers believe (or believed) that they were too smart and professional to need a union. And of course, if management gets wind of it while you are working on setting up the vote, they may do all kinds of trickery, legal or illegal, to block it.
It is possible. But it's quite uncommon in this industry in the US.
And given the current administration, it is hard to trust you'd get fair enforcement of labor laws if the company did illegal things to block unionization.
>The union demanded clear protections to ensure that recordings of actors’ performances could not be copied without consent and compensation.
>California’s legislation passed AB 2602 and AB 1836 in September 2024, prohibiting media companies from using AI to replicate actors’ performances without their consent.
it's a very reasonable law, but it is not the win you seem to believe it to be. the actors who refuse to consent will simply be passed over in favor of those who don't. no law will ever be passed to force companies to employ humans over machines, and if it were, the industry would move elsewhere. this is not without precedent :)
speech models have got so good so quickly that you can already replace a VA -- even an AAA prima donna -- with a teenager from Fiverr, who will simply bruteforce the right inflection.
I don't think software developers need to unionize - but to form French style co-ops.
e.g a lot of video game studios even the AAA ones could be co-ops. same as a lot of SAAS software companies. Linear - just announced a tender offer. I don't see a reason - why linear couldn't work as a co-op.
Because software developers enjoyed scarcity for most of the fields existence and could barter themselves better conditions instead of falling back to unionized fixed incomes.
In any case, the overwhelming majority of jobs is non-unionized and I don't think that software developers can stop technological changes by unionizing.
My objection has always been that advocates either can't or don't care to explain how such a thing would actually work from the perspective of a unionized engineer.
A lot of it will get automated the same way very many industries got automated. A lot of physical labour got automated once a primitive for it was created. Similarly, we have now a primitive for automating knowledge work. In the next few years to a decade, as all the right training data and runtime environments are slowly consolidated for various fields, a lot will be automated. There is no inherent reason a voice actor must be an eternal job, the same way there was no inherent reason for a draftsman or stage musician to be an eternal job.
It has nothing to do with skill. Both were very skilled jobs. Draftsman as well despite many people going to it straight from school. But computers and CAD mean that it is now necessary for someone to do a STEM degree to be a draftsman. Recorded audio made many stage musicians redundant. It is cheaper to do it this way and gets superior results, that is all, there is no further agenda.
Now too, the next generation of voice actors and many other knowledge workers will have to go up the value chain one step and operate or potentially build these tools (in whatever form they mature to in a decades time).
The current generation of voice actors will face the same situation as many before in the performance industry - stage musicians/performers for example that were made redundant by recorded audio. The reality is that most of them just left and dispersed into the economy doing completely unrelated jobs.
For software and generally computer engineers, this new primitive happens to itself be software, so it's less of a transition and an easier upskilling path to learn to build it. And building it is one step higher in the value chain than simply using it. That is a structural advantage.
I don't really like comparing the way automation of the past displaced jobs to the way AI is/will displace jobs. The timeline is just faster and, more importantly, there were still plenty of other fields of work for people to go to.
But now that we are automating white collar work... where will people go? I'm a 36 year-old veteran who has returned to college and so many of the younger students seem to be filled with despair.
I have hope for the future, but I think there will be an uncomfortable period of time.
I think people are overestimating the speed at which the transition will happen. I think it will happen much slower than is often portrayed online.
This slow transition period will also help answer the "where will they go" questions. We can't answer them right now.
Ultimately, everything we do is in service of social political and personal human incentives, and I think the effect of that is discounted when people make these takeoff predictions for AI and "AGI"
I'm really sad about how little creative control these tools have. They seem great for creating slop, but pretty useless for creating content that someone would love. I love the potential but text isn't really a great medium for describing artistic vision.
All these demos are focused on how easy it makes everything. Easy is great, but if everyone is able to make instant cute cat videos or whatever it just devalues it. I want to see turning a photo into a rigged 3d model, letting the artist animate and then generate the video. This technology could be used to increase creative expression, but instead it's being used to squeeze out creative expression
Text is not the only input. You can provide 3d block outs with rudimentary animation, annotated images with arrows etc, voice recordings of one person acting out some emotion then mapping that to a different character's voice, other uses of video to video, etc.
There could easily be at least some time period of low skilled ugly people acting in approximate but shitty ways in cheap sets just to give an input reference to a model and then describing the differences in text, yielding gorgeous people speaking with prestigious accents doing stuff in fancy locations in the output.
It'll come quickly for audio books. I've been working on a locally hosted, fully containerized web application to narrate my sci-fi novel using a full cast of characters and a few distinct narrators:
I would pay a few actors (mostly Star Trek TNG cast) money for a license to their voice.
I don't know how the next generation of beloved actors comes about and how we don't descend into a pit of neverending photocopies of things people once loved in the 1990s/2000s.
* $250 per finished hour for the narrator (lowest professional rate).
* 200k to 270k words.
* 22 hours, 20 hours, and 25 hours.
* Books 1 to 3 cost $5,500, $5,000, and $6,250, respectively.
My novel has 8 major characters, including the 3 narrators, and 25+ minor characters. That price tag is daunting, presumably in USD. This would mean investing $7,750 CAD in crafting an audiobook, which may not even sell, much less recoup the investment.
In contrast, a locally hosted solution has a wallet cost of pennies for electricity plus my time to develop the system.
>Firefox makes up about 90 percent of Mozilla’s revenue, according to Muhlheim, the finance chief for the organization’s for-profit arm — which in turn helps fund the nonprofit Mozilla Foundation. About 85 percent of that revenue comes from its deal with Google, he added.
What people usually say is that Google merely wants Firefox to survive for anti-competitive reasons. Presumably that does not necessitate it actually being used (or be usable).
Hot take: Google keeps Mozilla/Firefox alive through the default search engine placement (which makes Mozilla millions each year) so they don't get designated as a monopoly with their browser.
Apple forcing users to use their own browser/browsing engine doesn't disprove my argument IMO, virtually nobody outside the apple ecosystem uses Safari, and outside the apple ecosystem is something between 80-90% of internet users.
Firefox really struggle with demo pages of text-to-video models because of the large numbers of videos in the page in my experience, this page seems to work quite fine for me tho.
So Seedance is good primarily because of TikTok and this because of YouTube. I wonder what portion of all recorded video is privately held in hard drives at people’s homes or Apple photos. Of course there is data labeling and cleaning but is the next evolution just a question of access? Same goes for LLMs. Would people be willing to sell their data? Kind of a messed up way to make yourself obsolete. Or there is a limit to scaling?
I think they also have the problem of having given away their pro subscription to 10s or 100s of millions of students worldwide. They're tightening down on that now, and I have a feeling that this goes into them not releasing a larger model.
They blew up my interest when the stole my money by cutting me off from Gemini CLI with no explanation or recourse. I did not violate the terms of service and my only crime seemed to be not wanting to use Antigravity. They still took my money for the rest of that month and gave me nothing for it.
Just because Anthropic and OpenAI really want there to be an arms race justifying the outsized investment, doesn't mean the optimal play is to build larger, more expensive, models.
The capital infusion the frontier labs have received has gotten to a size where many believe it may not be possible to recoup this investment without some very unrealistic things happening.
I think it's reasonable to not completely drain one's cash reserves trying to stay ahead in a race where participants may very clearly be about to run straight off of a cliff.
If the Chinese labs can compete on a shoestring budget with access to much less powerful hardware, Google should be able to compete as well. They're becoming almost irrelevant for agentic coding right now.
It's not much of a shoestring budget to be receiving regular injections of investment from state lenders along with cheap credit.
I don't think the comparison holds.
And it doesn't have to be either/or. They could make larger, more expensive models, just at a slower cadence.
Sure downside would be not learning from people using your model for coding, if we're on the cusp of huge leaps in self-improvement. But there is a reasonable case for avoiding desperate scramble, especially if other parts of the business can also create value with the compute.
I work there. I have zero internal knowledge about the model. Opinion my own, etc. I don't think it is worth fighting to win on a month to month time horizon. When you step back and look an inch above this market, Gemini Pro 3.1 as a product was released in February. 6 months. It feels like forever and that Google is behind, but on a 2-3 year horizon? The models are going to stay similar.
Also, look at Flash 3.5 to 3.7. Flash 3.7 is a genuinely decent Sonnet 5 class model. Flash 3.7 is quite efficient too. Also, whatever was spent training 3.5 pro is probably not wasted. However, as a strategy, when I see models like Kimi K3, Fable, Sol. If you discard "because the model sucked" what other alternatives or potential options might exist?
I thought of a quite a few and they are far more compelling and interesting to me.
(Also Gemini models tend to be pretty decent at more than just programming. Enterprise AI use is more than just software eng / programming)
I'm the CTO of a GCP shop with an 8 figure annual commit.
If you'd told me at the end of Cloud Next 2025 that by now Google still wouldn't have a competitive offering to agentic coding offerings from Anthropic (Claude Code + Fable) or OpenAI (Codex + Sol), I wouldn't have believed you.
In our non-coding use cases where we're embedding models in our product, we're also not reaching for GCP stuff. Because Anthropic has the mindshare of our engineers and product folks, since it's what they use every day.
3.5 was almost certainly a 3.1 post-train, so likely a small investment on Google's part.
They mentioned that they have already started pretraining Gemini 4, which will be the full ground up rip-your-face-off-expensive training that is often discussed.
Yes, I agree with you that the race all the AI companies are running doesn't make sense, but at the same time, there are rumors that Google has produced newer versions of Pro without releasing them to the public.
Version 3.1 has plenty of room for improvement, yet they don't seem to be giving the attention it deserves or at least communicating accordingly.
There is more to the cost of a model than its training.
While training is a significant Capex expenditure, it has very low Operational cost after training unless it is deployed for public inference.
It may be that they wish to slow their cadence of releases, or develop their models to focus more in a different direction, etc. No matter what the actual reasoning, they have chosen to not compete in the same race, and I cannot say I fault them.
Google doesn't have a good coding model. This is a HUGE problem. They don't need "larger more expensive models", they need a good coding model because it's a competitive advantage.
I have a pro subscription, I think they have just given up. Likely because when they test their new models against the other frontier models they are so bad, they just pull it back. This leads them to try and innovate in other areas where there is currently less competition so they can compete. Not a bad play.
In the real world out there, Google and Microsoft are absolutely dominating enterprise customers.
Every single non-tech office worker I know is writing Gemini "gems" (sort of claude prompts/skills) or prompting Copilot to help drafting board meeting notes, insurance contracts updates that reflect changes in regulations, make quick loan feasibility assessments before passing them to the relevant office, presentations, etc, etc.
I'm talking insurance, banking, consultancy, manufacturing, etc, etc.
Why? Because Google and Microsoft already were in these companies, all they had to do is "oh, you also have AI now with your plans". Procurement and data compliance are the first thing businesses have to sort out. They were already sorted out.
Google doesn't need to have the best coding model or triumph in meaningless benchmarks, it only needs their models to get better and cheaper while serving them to their existing customer base.
They are playing a different game.
And Microsoft, doesn't even need to care about models at all, they can provide whatever open or closed AI with their services and have to focus on the harness in Excel or Github/Azure Copilot or whatever.
E.g. while developers in most of my clients use whatever they prefer or the company pays for, the remaining 90% uses either Google or Microsoft products.
Not a single one has incentives into venturing into OpenAI or Anthropic or Z.Ai lands because they might be better at some benchmark that is completely irrelevant to their tasks of updating powerpoints or summarizing incoming emails.
And, Youtube is huge both as a place where video contents goes and where can be trained from. Microdramas are starting to become a real category--14 Billion USD, 90% of it made with AI.
Chinese video models can be more immediately impressive, but none of them come close to beat the value of Google's Flow. Especially when you are throwing away a lot of generations as part of the creative process. Which is what you have to do to make longer content with any video model.
OpenAI needed to be able to focus. Google can walk and chew gum, and they're not going to run out of money to buy chewing gum.
Previously when making video ads you'd need to actually create the video. Actors, cameramen, editors - you name it. Now a new video ads is just a prompt away, directly inside the ad-spend web UI too no doubt.
People say Google have lost and that they're having their lunch eaten by anthropic, but I am not so sure...
Google owns 15% of Anthropic, Claude trains and runs on TPUs, and Google cloud is backlogged with demand from both OAI and Anthropic.
Google is selling shovels, leasing mines, buy stakes in "competitors" and doing it's own exploration/mining. When you look at the full picture, it kinda doesn't even look like Gemini matters that much to them overall.
They're main edge has been multimodal. I think they're still the best overall on multimodal? If I were them I would try to be the best at at least something.
I never personally got the least bit excited about Sora or nanobanana or whatever video/audio generation thing. But I guess I'm just not their customer. I do love the read-side of it though.
Sora was a social network type thing. Google sells their models on a PAYG basis - and makes money off them. Nano Banana alone has changed advertising 2D mockups and Photoshop like tasks forever. Notice how GPT image 2 is now available also on a PAYG basis.
I work in advertising and some days I spent a lot of money using these models. The amount and rapidity of prototyping using them has changed everything about advertising pre production.
The AI brand fragmentation at Google is not yet a problem because everyone is pretending:
x There are so-called “SOTA” or “frontier” models that are more effective than the other ones (independent of harnessing and routing)
x OpenAI and Anthropic have all the SOTA models and lead all the innovation
x Google’s moat is its search bread/butter (it’s the only reason they’re relevant)
All 3 operating assumptions are - I think - false.
What Google has done that the “cuter products” (Claude, ChatGPT) haven’t is connected relatively standard LLMs to an externally valuable live service.
As more companies realize that is where all the value is (the service) and not in the AI capability, then products (and humans) become important again.
Google should just be Google again, and Gemini should be Gemini, off to the side. Omni confuses everyone (and angers some iykyk), they should resolve “AI mode”, rename Gemma? and consolidate the brand overall so it’s clear what Google is.
Google is search.
It helps people on all sides of the market find what they’re looking for.
I don’t really see how repeatedly reinventing and rebranding the same AI chat UX is accomplishing anything toward that goal.
Implicit to that is "find". Their AI integration into search has really hit its stride for me. They have that search box (or speech prompt) hard wired into people and they are finally iterating and crafting AI into that experience. They really failed hard initially.
I know others have worse experiences than me but Google knows a lot about me so maybe that affects my results. YMMV
It does. Very commonly it will get acronyms wrong or assume I mean something else, even when it should be in my “ad profile” or whatever it bases it off.
Often I think the search engine is working perfectly then the LLM is ruining it in delivery.
I'm still getting major uncanny valley from any of the videos featuring humans, something about them disgusts me. I guess I should be glad I'm still able to distinguish them.
Something in their eyes. Looks very robotic / lifeless for me. And the sound-mixing is very off. Clearly feels like the voice was layered on top of whatever sound is in the background and not blended.
I don't think I can see the difference. I just have my skin crawl because I'm expecting to see something off and generally have a bad feeling about it.
It certainly makes for easy demos, but I always struggle with the practical application. As in, what work or enjoyment does someone actually get from this? Ads and media pre production seem plausible, but it fails the 'how can this enrich life' in a way most other AI tools don't. Maybe for them that's not a consideration, if their only interest is the other meaning of enrich that might flow from ads and numbing rivers of slop.
Why do we look at art, watch videos/movies? Is that replicable as a function of text, other existing media, and 3-30 cents of compute per second? I'm pretty functionalist about these things, and at some point it probably won't be possible to tell the difference. But until then, at which point we might just say 'death of the author', it seems like a category error.
I do work with artists that use video and image generation models to create stuff, but from what I can tell they're interested in faster iteration and controlling a lot of intermediate steps (their graphs can get pretty labyrinthine).
Yes, most people using AI for creative projects spend a lot of time and attention mastering their tools, and figure out how to adapt them into their creative processes. AI can dramatically lower the cost of indie productions, while also a allowing a broader range of stories to be told. Even the most successful film makers need to bow and scrape to get their projects funded, democratizing visual media can be a good thing, even if you, personally, are no more likely to do this than you are to pick up Photoshop or record a podcast.
I enjoy making short films with AI. When my latest short screens at a festival in Ocotber, alongside traditional and AI films, hopefully the audience will like it too.
The quick "one shot" video generation might be slop to you, or I. But if someone wants to send it as birthday greeting to their aunt, and they both enjoy it, what business is it of ours?
Kids love image and video generation! The former is cheap enough to do just because it is fun.
I have a young boy, and whenever he builds an impressive "scene" from LEGO (like a diorama or whatever), I take a couple of reference pictures with my phone and make it into a "real" movie scene, cartoon, or whatever. He loves it, and this motivates him to build more and bigger things out of LEGO.
If he builds something really special, I might actually fork over the $5 to use Omni to turn his LEGO creation into a 10-second video instead of a still image. It'll blow his mind!
PS: There also are cheap and even free phone apps that make stop-motion animation trivial. We've already made a couple of videos of his toys moving around that way.
I let myself get mildly excited with the last Omni release, but it turns out it (and this one) can't do the one practical thing I want - Sync generated video to provided pre-existing audio.
Meanwhile, I'm happily using Minimax H3 locally on my 12Gb 4070RTX to finally finish the lip syncing to recorded dialog on my abandoned 20 year old Flash animation hobby projects.
Pretty sure I heard one of the PMs in a podcast a few weeks ago say they are intentionally not building support for it out of concerns of enabling deep-fakes.
I agree though. My issue is the cost for using AI video models is way too high for anyone not building anything serious with them, at the same time they are too restricted for actually using professionally. Prompting them with text to get something generated is cute, but then you just end up creating slop that everyone hates, ultimately devaluing the power of these things.
While it sounds great you're quickly disappointed after you run the same prompt at standard resolution only to get a different result because it's non deterministic.
Generate lightweight previews in 360p resolution up to 60% faster
and at a third of the cost compared to Omni 1.1’s standard 720p resolution. This is helpful for rapid prototyping, storyboard iteration, and quick rendering in developer platforms."
This is a great idea, to have a low-resolution mode for additional speed to create previews, do test runs, create rapid prototypes, etc.
My curiousity is, what's the absolute useable minimum that this could be?
That is, would/could 240p resolution work? If so, what about 144p? How about lower? Then, could those images be upscaled quickly (and is the result still usable?) with a faster image upscaling-only neural network?
The reason why knowing such lower numbers / lower bounds -- is because they could be important for additional cost/time savings and/or running derived LLM's on local resource-constrained hardware...
Anyway, great post, great idea, and we welcome Gemnini Omni 1.1 Flash to the ever-expanding list of LLM/AI's!
petcat | 10 hours ago
newyankee | 10 hours ago
lambda | 10 hours ago
For example: https://sites.suffolk.edu/jhtl/2025/10/30/game-over-for-unau...
Software developers, unfortunately, have been convinced that they don't need to unionize, so have no collective bargaining power for dealing with situations like this.
r_lee | 10 hours ago
I think too many just expect that they can get by like before because they're a 10x engineer or whatever
quaintdev | 10 hours ago
georgemcbay | 9 hours ago
In my experience this is what almost everyone thinks, right up to the point where it starts happening to them.
And it isn't just an individual blind spot, organizations suffer from the same thing in the sense that many companies will be happy to automate away all their labor to avoid paying the "human tax" without giving much thought to the fact that the need for the company itself will also be automated away soon after.
If you can replace nearly all of your workers with some cheap tokens and prompts, everyone else who previously would have been a customer can replace your whole company with the same thing.
embedding-shape | 10 hours ago
To be fair, most of us spent our entire career trying our best to automate ourselves away one way or another, and always seen that as our job description.
12ags-18276 | 10 hours ago
Some people dislike automation altogether and like the computational aspect. Most people seem to enjoy constructing virtual worlds and Rube Goldberg machines.
genericone | 8 hours ago
12ags-18276 | 10 hours ago
Useful idiots always go first after the goal is accomplished.
zamadatix | 9 hours ago
Inexplicably for the current moment, AI has so far actually meant I have even more to do instead of less. I expect that to change eventually... but man, is it a bit of whiplash to go from wondering if my job will be around in 10 years to starting the work day and being more backed up than ever. Doubly so since the rate of change does not seem to be very evenly distributed by tech role.
oblio | 8 hours ago
Spoiler alert: unions can help with that.
zamadatix | 7 hours ago
BeetleB | 9 hours ago
It's kind of hypocritical to change just because it's now a different industry being impacted.
qwytw | 8 hours ago
GeorgeWBasic | 6 hours ago
catigula | 10 hours ago
https://www.computerweekly.com/news/252481961/Amazons-Whole-...
12ahsgf | 9 hours ago
https://www.sciencedirect.com/science/article/abs/pii/S01762...
Lagoon3384_SWE | 10 hours ago
Sweden just has an "office worker" union.
https://www.unionen.se/in-english/this-is-unionen
Now, they cannot prevent offshoring or AI use or replacements. Very very rarely if ever.
But they can negotiate a large layoff and give better pay/conditions during work etc. This view that unions can prevent anything... when so many jobs have gone to east EU (all amazing and good at what they do but lets face it, it is due to pay) so yea there wont be any union preventing AI...
arjie | 9 hours ago
That’ll be somewhat detrimental to me perhaps (which isn’t certain), but that’s okay. I’d rather attempt to adapt to a changing world.
koe123 | 9 hours ago
smallerize | 9 hours ago
BeetleB | 9 hours ago
smallerize | 9 hours ago
BeetleB | 9 hours ago
It is if it makes it worse for everyone else.
lambda | 9 hours ago
A union doesn't give you unilateral power; it just gives you a better seat at the bargaining table. Capital pools its resources to negotiate better as a single bloc; why shouldn't labor as well?
iAMkenough | 8 hours ago
Yes, exactly. Do not advocate for yourself and your colleagues by joining them in a united front. Your voice does not matter. You will continue to adapt to widening wage disparities.
You will accept a pay cut and increased productivity goals and be happy to continue adapting as your cost of living keeps on climbing.
brendoelfrendo | 7 hours ago
jackdoe | 7 hours ago
localized gains, distributed loss.
computerliker | 7 hours ago
a computer can never be sad or horny or have taste. it can only make product. i also think artists should be paid living wages with good working conditions.
thank you SAG-AFTRA, WGA, IATSE, AEMI and Teamsters.
binary132 | 3 hours ago
sib | 7 hours ago
Unions are very much an in-group vs out-group phenomenon (and in many cases, the benefits are specifically to the more senior union members vis a vis the less senior ones.)
magicalist | 4 hours ago
If you're talking about AB 1751 (Missing Middle Townhome Ownership Act) I think you have this exactly backwards.
> Unions are very much an in-group vs out-group phenomenon
As opposed to what? You and the GGP phrase this as a dichotomy, but I'm really curious what the other side of this is, because "not being in a union" doesn't erase your self interest or all the many many groups of people doing the same that don't feel like they're apparently being overly self interested by forming corporations, governments, non-profits, advocacy groups, etc, etc, etc
baron816 | 6 hours ago
mattkevan | 6 hours ago
baron816 | 4 hours ago
Unions can be a deterrent to this. If you have sectoral bargaining, you can’t get a better deal anywhere because everyone pays the same. Or even if it’s within a single company, you’re just stuck with whatever scheme the union bosses agreed to.
But also, unions are a deterrent to new employers since fewer people want to start companies in industries where unions have taken over. They’re a nightmare to deal with. Fewer employers means less competition means worse services.
We have a dynamic economy with lots of competition. It’s easy to switch jobs and find a better place to work. We should not want to give this up.
mchusma | 2 hours ago
I don’t have any objection to voluntary unions. Meaning people can join them for support or collective bargaining, but that citizens can just opt to not employ anyone in the union. Unions could be support structures for workers, instead they are mafia bosses.
lambdas | 5 hours ago
baron816 | 4 hours ago
itishappy | 5 hours ago
csallen | 4 hours ago
I read it more like they're against trying to stop, hinder, or prevent the use of new technology, simply because it's personally threatening.
And I feel the same way. If I was part of a union, I would believe in the collective bargaining power, but I wouldn't necessarily believe we should use that to pursue every single possible agenda that might benefit us. And I've personally pooled my capital with others in the past in order to achieve certain outcomes. And in those situations I also didn't attempt to reach so far as to limit the totally valid freedoms of others just because I knew it might be better for me.
Acting in the interest of a collective doesn't mean one must also abandon the ethics they usually employ when acting in their solo self interest.
jplusequalt | 6 hours ago
Caring about your ability to pay your bills, and support your family is "self-interested" in the same way that choosing to not drive into oncoming traffic is "self-interested".
popalchemist | 2 hours ago
timr | an hour ago
The former is obviously unfair to the person whose identity is stolen. The latter is sour grapes from people who want to control the future.
I also can't bring myself to get incensed over someone using a robot to write code. It's literally the modern equivalent of smashing knitting looms.
boredatoms | 9 hours ago
lambda | 7 hours ago
It is possible. But it's quite uncommon in this industry in the US.
And given the current administration, it is hard to trust you'd get fair enforcement of labor laws if the company did illegal things to block unionization.
vlyan | 9 hours ago
>California’s legislation passed AB 2602 and AB 1836 in September 2024, prohibiting media companies from using AI to replicate actors’ performances without their consent.
it's a very reasonable law, but it is not the win you seem to believe it to be. the actors who refuse to consent will simply be passed over in favor of those who don't. no law will ever be passed to force companies to employ humans over machines, and if it were, the industry would move elsewhere. this is not without precedent :)
speech models have got so good so quickly that you can already replace a VA -- even an AAA prima donna -- with a teenager from Fiverr, who will simply bruteforce the right inflection.
isoprophlex | 8 hours ago
worthless-trash | 2 minutes ago
dzonga | 7 hours ago
e.g a lot of video game studios even the AAA ones could be co-ops. same as a lot of SAAS software companies. Linear - just announced a tender offer. I don't see a reason - why linear couldn't work as a co-op.
ameliaquining | 2 hours ago
epolanski | 7 hours ago
In any case, the overwhelming majority of jobs is non-unionized and I don't think that software developers can stop technological changes by unionizing.
ameliaquining | 2 hours ago
earthnail | 10 hours ago
It’s brutal for these people. The creative industry was always hard, but this is just plain brutal.
moscoe | 9 hours ago
tiagod | 8 hours ago
chung8123 | 3 hours ago
oblio | 8 hours ago
porridgeraisin | 9 hours ago
It has nothing to do with skill. Both were very skilled jobs. Draftsman as well despite many people going to it straight from school. But computers and CAD mean that it is now necessary for someone to do a STEM degree to be a draftsman. Recorded audio made many stage musicians redundant. It is cheaper to do it this way and gets superior results, that is all, there is no further agenda.
Now too, the next generation of voice actors and many other knowledge workers will have to go up the value chain one step and operate or potentially build these tools (in whatever form they mature to in a decades time).
The current generation of voice actors will face the same situation as many before in the performance industry - stage musicians/performers for example that were made redundant by recorded audio. The reality is that most of them just left and dispersed into the economy doing completely unrelated jobs.
For software and generally computer engineers, this new primitive happens to itself be software, so it's less of a transition and an easier upskilling path to learn to build it. And building it is one step higher in the value chain than simply using it. That is a structural advantage.
bubblemoth | 9 hours ago
But now that we are automating white collar work... where will people go? I'm a 36 year-old veteran who has returned to college and so many of the younger students seem to be filled with despair.
I have hope for the future, but I think there will be an uncomfortable period of time.
porridgeraisin | 8 hours ago
This slow transition period will also help answer the "where will they go" questions. We can't answer them right now.
Ultimately, everything we do is in service of social political and personal human incentives, and I think the effect of that is discounted when people make these takeoff predictions for AI and "AGI"
space_fountain | 9 hours ago
All these demos are focused on how easy it makes everything. Easy is great, but if everyone is able to make instant cute cat videos or whatever it just devalues it. I want to see turning a photo into a rigged 3d model, letting the artist animate and then generate the video. This technology could be used to increase creative expression, but instead it's being used to squeeze out creative expression
bonoboTP | 7 hours ago
There could easily be at least some time period of low skilled ugly people acting in approximate but shitty ways in cheap sets just to give an input reference to a model and then describing the differences in text, yielding gorgeous people speaking with prestigious accents doing stuff in fancy locations in the output.
NooneAtAll3 | 8 hours ago
one youtuber used irl footage, with hands and stuff - so I know there's human behind the camera
the other was a letsplay that reacted to events just fine emotionally
and yet the uncanny valley of the sound is in full force. Maybe youtube has done something with the codecs?
I feel sad
thangalin | 8 hours ago
* https://i.ibb.co/ccqKZ71L/keenlore.png
* https://i.ibb.co/1t3W0JqZ/keenlore-02.png
* https://i.ibb.co/LdBqHwKB/keenlore-03.png
dormento | 7 hours ago
ianmarcinkowski | 7 hours ago
I don't know how the next generation of beloved actors comes about and how we don't descend into a pit of neverending photocopies of things people once loved in the 1990s/2000s.
thangalin | 6 hours ago
* $250 per finished hour for the narrator (lowest professional rate).
* 200k to 270k words.
* 22 hours, 20 hours, and 25 hours.
* Books 1 to 3 cost $5,500, $5,000, and $6,250, respectively.
My novel has 8 major characters, including the 3 narrators, and 25+ minor characters. That price tag is daunting, presumably in USD. This would mean investing $7,750 CAD in crafting an audiobook, which may not even sell, much less recoup the investment.
In contrast, a locally hosted solution has a wallet cost of pennies for electricity plus my time to develop the system.
ChickeNES | 5 hours ago
To me it's been obvious for a while now: there won't be one
unified101 | 3 hours ago
dominotw | 8 hours ago
it always sucks at everything else.
037 | 10 hours ago
fg137 | 10 hours ago
fl4regun | 9 hours ago
>Firefox makes up about 90 percent of Mozilla’s revenue, according to Muhlheim, the finance chief for the organization’s for-profit arm — which in turn helps fund the nonprofit Mozilla Foundation. About 85 percent of that revenue comes from its deal with Google, he added.
https://www.theverge.com/news/660548/firefox-google-search-r...
jorl17 | 3 hours ago
sva_ | 9 hours ago
BbzzbB | 9 hours ago
Vinnl | 8 hours ago
oblio | 8 hours ago
Billion with a B, as in 10 zeroes, not 7.
sva_ | 7 hours ago
static_motion | 8 hours ago
WarmWash | 8 hours ago
If anything FF gets left out because usage is so low.
oblio | 8 hours ago
I've been following this for a long time. They were leaving Firefox out when its usage wasn't low.
Probably the typical backdoor executive mandate that led to death by "sprint prioritization":
Yes, we will for sure work on the Firefox compatibility bug, Dave-Open-Source-Enthusiast-Google-Dev.
But we can only pick up 10 bugfixing tickets this sprint and the ticket you highlighted, as the entire team agrees, is priority #12.
<repeat every sprint, where during the sprint 9-10 new higher priority items magically appear just in time for the next sprint>
Death by slow asphyxiation.
mattlondon | 8 hours ago
tvst | 7 hours ago
mattlondon | 7 hours ago
GaggiX | 10 hours ago
ygouzerh | an hour ago
Gecko4072 | 10 hours ago
guilhermeasper | 10 hours ago
AISnakeOil | 9 hours ago
Pro models are mainly for coding agent work; it doesn't necessarily make them any money.
sva_ | 9 hours ago
jimmoores | 9 hours ago
SimianSci | 9 hours ago
The capital infusion the frontier labs have received has gotten to a size where many believe it may not be possible to recoup this investment without some very unrealistic things happening.
I think it's reasonable to not completely drain one's cash reserves trying to stay ahead in a race where participants may very clearly be about to run straight off of a cliff.
jimmoores | 9 hours ago
bitexploder | 8 hours ago
SimianSci | 8 hours ago
msabalau | 8 hours ago
Sure downside would be not learning from people using your model for coding, if we're on the cusp of huge leaps in self-improvement. But there is a reasonable case for avoiding desperate scramble, especially if other parts of the business can also create value with the compute.
onlyrealcuzzo | 8 hours ago
They never gave an official answer as to why, so I'll let you draw your own conclusions.
They did not decide it wasn't worth spending the money to train.
They absolutely spent the money.
bitexploder | 8 hours ago
Also, look at Flash 3.5 to 3.7. Flash 3.7 is a genuinely decent Sonnet 5 class model. Flash 3.7 is quite efficient too. Also, whatever was spent training 3.5 pro is probably not wasted. However, as a strategy, when I see models like Kimi K3, Fable, Sol. If you discard "because the model sucked" what other alternatives or potential options might exist?
I thought of a quite a few and they are far more compelling and interesting to me.
(Also Gemini models tend to be pretty decent at more than just programming. Enterprise AI use is more than just software eng / programming)
mh- | 7 hours ago
If you'd told me at the end of Cloud Next 2025 that by now Google still wouldn't have a competitive offering to agentic coding offerings from Anthropic (Claude Code + Fable) or OpenAI (Codex + Sol), I wouldn't have believed you.
In our non-coding use cases where we're embedding models in our product, we're also not reaching for GCP stuff. Because Anthropic has the mindshare of our engineers and product folks, since it's what they use every day.
WarmWash | 7 hours ago
They mentioned that they have already started pretraining Gemini 4, which will be the full ground up rip-your-face-off-expensive training that is often discussed.
guilhermeasper | 8 hours ago
Version 3.1 has plenty of room for improvement, yet they don't seem to be giving the attention it deserves or at least communicating accordingly.
SimianSci | 8 hours ago
It may be that they wish to slow their cadence of releases, or develop their models to focus more in a different direction, etc. No matter what the actual reasoning, they have chosen to not compete in the same race, and I cannot say I fault them.
gbriel | 7 hours ago
dbbk | 6 hours ago
jnwatson | 5 hours ago
dbbk | 6 hours ago
mianos | 7 hours ago
epolanski | 7 hours ago
In the real world out there, Google and Microsoft are absolutely dominating enterprise customers.
Every single non-tech office worker I know is writing Gemini "gems" (sort of claude prompts/skills) or prompting Copilot to help drafting board meeting notes, insurance contracts updates that reflect changes in regulations, make quick loan feasibility assessments before passing them to the relevant office, presentations, etc, etc.
I'm talking insurance, banking, consultancy, manufacturing, etc, etc.
Why? Because Google and Microsoft already were in these companies, all they had to do is "oh, you also have AI now with your plans". Procurement and data compliance are the first thing businesses have to sort out. They were already sorted out.
Google doesn't need to have the best coding model or triumph in meaningless benchmarks, it only needs their models to get better and cheaper while serving them to their existing customer base.
They are playing a different game.
And Microsoft, doesn't even need to care about models at all, they can provide whatever open or closed AI with their services and have to focus on the harness in Excel or Github/Azure Copilot or whatever.
E.g. while developers in most of my clients use whatever they prefer or the company pays for, the remaining 90% uses either Google or Microsoft products.
Not a single one has incentives into venturing into OpenAI or Anthropic or Z.Ai lands because they might be better at some benchmark that is completely irrelevant to their tasks of updating powerpoints or summarizing incoming emails.
vismit2000 | 28 minutes ago
simonw | 9 hours ago
Maybe because they see video generation as key to developing "world models"?
kfarr | 9 hours ago
OtherShrezzing | 9 hours ago
bahmboo | 8 hours ago
msabalau | 8 hours ago
And, Veo and omni simply were better than Sora
And, Youtube is huge both as a place where video contents goes and where can be trained from. Microdramas are starting to become a real category--14 Billion USD, 90% of it made with AI.
Chinese video models can be more immediately impressive, but none of them come close to beat the value of Google's Flow. Especially when you are throwing away a lot of generations as part of the creative process. Which is what you have to do to make longer content with any video model.
OpenAI needed to be able to focus. Google can walk and chew gum, and they're not going to run out of money to buy chewing gum.
pwython | 5 hours ago
https://youtu.be/QaiecWzeHFM?si=UV6eF-m514nCOFe0&t=156
mattlondon | 8 hours ago
Previously when making video ads you'd need to actually create the video. Actors, cameramen, editors - you name it. Now a new video ads is just a prompt away, directly inside the ad-spend web UI too no doubt.
People say Google have lost and that they're having their lunch eaten by anthropic, but I am not so sure...
WarmWash | 7 hours ago
Google is selling shovels, leasing mines, buy stakes in "competitors" and doing it's own exploration/mining. When you look at the full picture, it kinda doesn't even look like Gemini matters that much to them overall.
simonw | 6 hours ago
chermi | 7 hours ago
I never personally got the least bit excited about Sora or nanobanana or whatever video/audio generation thing. But I guess I'm just not their customer. I do love the read-side of it though.
kranke155 | 7 hours ago
I work in advertising and some days I spent a lot of money using these models. The amount and rapidity of prototyping using them has changed everything about advertising pre production.
tiahura | 9 hours ago
CmonGoogle | 9 hours ago
x There are so-called “SOTA” or “frontier” models that are more effective than the other ones (independent of harnessing and routing)
x OpenAI and Anthropic have all the SOTA models and lead all the innovation
x Google’s moat is its search bread/butter (it’s the only reason they’re relevant)
All 3 operating assumptions are - I think - false.
What Google has done that the “cuter products” (Claude, ChatGPT) haven’t is connected relatively standard LLMs to an externally valuable live service.
As more companies realize that is where all the value is (the service) and not in the AI capability, then products (and humans) become important again.
Google should just be Google again, and Gemini should be Gemini, off to the side. Omni confuses everyone (and angers some iykyk), they should resolve “AI mode”, rename Gemma? and consolidate the brand overall so it’s clear what Google is.
Google is search.
It helps people on all sides of the market find what they’re looking for.
I don’t really see how repeatedly reinventing and rebranding the same AI chat UX is accomplishing anything toward that goal.
bahmboo | 8 hours ago
Implicit to that is "find". Their AI integration into search has really hit its stride for me. They have that search box (or speech prompt) hard wired into people and they are finally iterating and crafting AI into that experience. They really failed hard initially.
I know others have worse experiences than me but Google knows a lot about me so maybe that affects my results. YMMV
CmonGoogle | 5 hours ago
It does. Very commonly it will get acronyms wrong or assume I mean something else, even when it should be in my “ad profile” or whatever it bases it off.
Often I think the search engine is working perfectly then the LLM is ruining it in delivery.
There are other UIs besides chat
polytely | 9 hours ago
aloknnikhil | 9 hours ago
kzrdude | 9 hours ago
bahmboo | 8 hours ago
When it comes to fish swimming around I don't think I would be able to reliably tell what was real vs generated even with deep inspection.
overflowy | 8 hours ago
visarga | 7 hours ago
rcr-anti | 8 hours ago
Why do we look at art, watch videos/movies? Is that replicable as a function of text, other existing media, and 3-30 cents of compute per second? I'm pretty functionalist about these things, and at some point it probably won't be possible to tell the difference. But until then, at which point we might just say 'death of the author', it seems like a category error.
I do work with artists that use video and image generation models to create stuff, but from what I can tell they're interested in faster iteration and controlling a lot of intermediate steps (their graphs can get pretty labyrinthine).
msabalau | 8 hours ago
I enjoy making short films with AI. When my latest short screens at a festival in Ocotber, alongside traditional and AI films, hopefully the audience will like it too.
The quick "one shot" video generation might be slop to you, or I. But if someone wants to send it as birthday greeting to their aunt, and they both enjoy it, what business is it of ours?
visarga | 7 hours ago
jiggawatts | 8 hours ago
I have a young boy, and whenever he builds an impressive "scene" from LEGO (like a diorama or whatever), I take a couple of reference pictures with my phone and make it into a "real" movie scene, cartoon, or whatever. He loves it, and this motivates him to build more and bigger things out of LEGO.
If he builds something really special, I might actually fork over the $5 to use Omni to turn his LEGO creation into a 10-second video instead of a still image. It'll blow his mind!
PS: There also are cheap and even free phone apps that make stop-motion animation trivial. We've already made a couple of videos of his toys moving around that way.
thisisauserid | 8 hours ago
visarga | 7 hours ago
I like to generate songs from obscure poems
Nihilartikel | 8 hours ago
Meanwhile, I'm happily using Minimax H3 locally on my 12Gb 4070RTX to finally finish the lip syncing to recorded dialog on my abandoned 20 year old Flash animation hobby projects.
msabalau | 8 hours ago
jarjoura | 4 hours ago
I agree though. My issue is the cost for using AI video models is way too high for anyone not building anything serious with them, at the same time they are too restricted for actually using professionally. Prompting them with text to get something generated is cute, but then you just end up creating slop that everyone hates, ultimately devaluing the power of these things.
doctorpangloss | 2 hours ago
SpyCoder77 | 8 hours ago
shreya1999 | 7 hours ago
chermi | 7 hours ago
cube00 | 7 hours ago
While it sounds great you're quickly disappointed after you run the same prompt at standard resolution only to get a different result because it's non deterministic.
kridsdale1 | 5 hours ago
cube00 | 5 hours ago
Their changelog suggests using their new upscaler but I've had nothing but disappointment from upscalers.
Upscale when ready: Users on paid tiers can seamlessly upgrade their favorite clips using our new 360p to 720p upscaler.
peter_d_sherman | 2 hours ago
Generate lightweight previews in 360p resolution up to 60% faster
and at a third of the cost compared to Omni 1.1’s standard 720p resolution. This is helpful for rapid prototyping, storyboard iteration, and quick rendering in developer platforms."
This is a great idea, to have a low-resolution mode for additional speed to create previews, do test runs, create rapid prototypes, etc.
My curiousity is, what's the absolute useable minimum that this could be?
That is, would/could 240p resolution work? If so, what about 144p? How about lower? Then, could those images be upscaled quickly (and is the result still usable?) with a faster image upscaling-only neural network?
The reason why knowing such lower numbers / lower bounds -- is because they could be important for additional cost/time savings and/or running derived LLM's on local resource-constrained hardware...
Anyway, great post, great idea, and we welcome Gemnini Omni 1.1 Flash to the ever-expanding list of LLM/AI's!