I heard a radio spot recently and I wondered if the voice was a real person or AI. It makes we wonder how such industries are dealing with this gen-AI revolution. We spend a lot of time here thinking about how it affects software developers, but I hardly ever see any commentary on how it is affecting screen and voice actors.
Many screen and voice actors are unionized, and the unions have been striking and bargaining specifcially over these points.<p>For example: <a href="https://sites.suffolk.edu/jhtl/2025/10/30/game-over-for-unauthorized-ai-performances-the-sag-aftra-video-game-strike-and-performer-protections-under-the-new-collective-bargaining-agreement/" rel="nofollow">https://sites.suffolk.edu/jhtl/2025/10/30/game-over-for-unau...</a><p>Software developers, unfortunately, have been convinced that they don't need to unionize, so have no collective bargaining power for dealing with situations like this.
Personally, I just can’t find it in me to be <i>that</i> self-interested. It’s how other people oppose buildings near them because it blocks their view. I really don’t want to stop other people from writing software. If they want to use AI to do it so be it.<p>That’ll be somewhat detrimental to me perhaps (which isn’t certain), but that’s okay. I’d rather attempt to adapt to a changing world.
I don't know that unions would necessarily negotiate for no use of AI. The SAG-AFRA deals don't preclude all use of AI; they just requre consent and negotiation in certain cases.<p>A union doesn't give you unilateral power; it just gives you a better seat at the bargaining table. Capital pools its resources to negotiate better as a single bloc; why shouldn't labor as well?
>Personally, I just can’t find it in me to be that self-interested<p>Caring about your ability to pay your bills, and support your family is "self-interested" in the same way that choosing to not drive into oncoming traffic is "self-interested".
This is the first time I've heard someone argue that unionizing is a self-interested pursuit. It's called "collective bargaining" for a reason; workers who cooperate in negotiations with their employers have more leverage than those who negotiate individually. The entire point is to work together for the betterment of the group!
See current situation where the Carpenters' Union in California is pushing hard against an initiative that would make it much easier to build more housing in the state because it would also reduce requirements for union labor.<p>Unions are very much an in-group vs out-group phenomenon (and in many cases, the benefits are specifically to the more senior union members vis a vis the less senior ones.)
Have you ever not been able to send your kids to school for weeks (or years in the case of COVID) because the teachers refused to work? Have you ever been unable to get timely medical treatment because nurses refused to work? Have you ever been unable to work yourself because the transit workers refused to work?
Maybe if employers offered better working conditions and salaries people wouldn’t feel the need to go on strike?
>Have you ever been unable to get timely medical treatment because nurses refused to work<p>My wife's a nurse. She works 12 hour shifts at $30 an hour to keep your grandparents alive while being berated by families, screamed at by drug addicts, hit by elderly patients, forced to wipe asses, etc. All the while she is liable if she fucks up and a patient decides to sue.<p>It's one of the hardest jobs anyone can work.<p>Go talk to some nurses and be in community with your fellow human before being overly self-centered, prick.
their point is while the 'group' wins, the rest of the world loses.<p>localized gains, distributed loss.
as an art enjoyer, i do not feel that having less media written, drawn, or voiced by a computer is a loss.<p>a computer can never be sad or horny or have taste. it can only make product. i also think artists should be paid living wages with good working conditions.<p>thank you SAG-AFTRA, WGA, IATSE, AEMI and Teamsters.
> Personally, I just can’t find it in me to be that self-interested.<p>Yes, exactly. Do not advocate for yourself and your colleagues by joining them in a united front. Your voice does not matter. You will continue to adapt to widening wage disparities.<p>You will accept a pay cut and increased productivity goals and be happy to continue adapting as your cost of living keeps on climbing.
You may be surprised to know that residential areas in San Francisco are height limited, sometimes as short as 4 stories. They're not exactly anti-progress.
Thats great, although understand that unionizing protects the more vulnerable of also the software engineered. You not advocating for your rights also means weakening others. Is it still self interested to unionized from that perspective?
>The union demanded clear protections to ensure that recordings of actors’ performances could not be copied without consent and compensation.<p>>California’s legislation passed AB 2602 and AB 1836 in September 2024, prohibiting media companies from using AI to replicate actors’ performances without their consent.<p>it's a very reasonable law, but it is not the win you seem to believe it to be. the actors who refuse to consent will simply be passed over in favor of those who don't. no law will ever be passed to force companies to employ humans over machines, and if it were, the industry would move elsewhere. this is not without precedent :)<p>speech models have got so good so quickly that you can already replace a VA -- even an AAA prima donna -- with a teenager from Fiverr, who will simply bruteforce the right inflection.
I don't think software developers need to unionize - but to form French style co-ops.<p>e.g a lot of video game studios even the AAA ones could be co-ops. same as a lot of SAAS software companies. Linear - just announced a tender offer. I don't see a reason - why linear couldn't work as a co-op.
it's hilarious to see some people here go on about how they're proud to work so that they can be replaced<p>I think too many just expect that they can get by like before because they're a 10x engineer or whatever
I think I sit somewhere between these two descriptions. I've always supported unions and other "for the common good" type machinery. At the same time, I desperately also don't want to end up doing something the equivalent of barring the use of calculators just so I can toil away at a 9-5 crunching numbers more slowly instead. If AI really does replace all of the meaningful jobs we can do... great - I'd rather we make sure the spoils of that production are distributed than try to cling to preventing the technology from being used.<p>Inexplicably for the current moment, AI has so far actually meant I have even more to do instead of less. I expect that to change eventually... but man, is it a bit of whiplash to go from wondering if my job will be around in 10 years to starting the work day and being more backed up than ever. Doubly so since the rate of change does not seem to be very evenly distributed by tech role.
> it's hilarious to see some people here go on about how they're proud to work so that they can be replaced<p>To be fair, most of us spent our entire career <i>trying our best</i> to automate ourselves away one way or another, and always seen that as our job description.
No, that is just nerd heroism lore that indeed has always shown up in comments, that part is definitely true.<p>Some people dislike automation altogether and like the computational aspect. Most people seem to enjoy constructing virtual worlds and Rube Goldberg machines.
I agree with your point, but I don't think it's funny at all. The whole field is changing because of this. The career, except for those ones lucky enough to have jobs where they're valued and they can decide how to use AI tools, is turning into shit because of this and software engineers have become a lot less valuable as management is trying to turn them into reverse centaurs.
Much of computing history has been about replacing humans. If they were proud of their work pre-LLMs, then at least they're being consistent!<p>It's kind of hypocritical to change just because it's now a different industry being impacted.
What options do you have? If you're the 10x engineer you'll be a 100x one and do just fine. Otherwise what can you do besides holding on as long as you can or start searching for a new career.
It won't happen to me is what I used to think
> It won't happen to me is what I used to think<p>In my experience this is what almost everyone thinks, right up to the point where it starts happening to them.<p>And it isn't just an individual blind spot, organizations suffer from the same thing in the sense that many companies will be happy to automate away all their labor to avoid paying the "human tax" without giving much thought to the fact that the need for the company itself will also be automated away soon after.<p>If you can replace nearly all of your workers with some cheap tokens and prompts, everyone else who previously would have been a customer can replace your whole company with the same thing.
Many people who promoted DEI when their corporations demanded it were fired later.<p>Useful idiots always go first after the goal is accomplished.
And thus the developers remain, until this very moment, ionized.
No one has offered a union to join at any workplace Ive been at
Yeah, the problem is, someone has to start the union. It takes work. You have to get the whole workforce to vote on it. A lot of software engineers believe (or believed) that they were too smart and professional to need a union. And of course, if management gets wind of it while you are working on setting up the vote, they may do all kinds of trickery, legal or illegal, to block it.<p>It is possible. But it's quite uncommon in this industry in the US.<p>And given the current administration, it is hard to trust you'd get fair enforcement of labor laws if the company did illegal things to block unionization.
Because software developers enjoyed scarcity for most of the fields existence and could barter themselves better conditions instead of falling back to unionized fixed incomes.<p>In any case, the overwhelming majority of jobs is non-unionized and I don't think that software developers can stop technological changes by unionizing.
[dead]
Half of them are from India or China. Racial diversity is known to be detrimental to unionization. In fact, Amazon used this exact strategy to bust nascent Somali unions.<p><a href="https://www.computerweekly.com/news/252481961/Amazons-Whole-Foods-uses-heat-mapping-to-track-unionisation-efforts" rel="nofollow">https://www.computerweekly.com/news/252481961/Amazons-Whole-...</a>
It'll come quickly for audio books. I've been working on a locally hosted, fully containerized web application to narrate my sci-fi novel using a full cast of characters and a few distinct narrators:<p>* <a href="https://i.ibb.co/ccqKZ71L/keenlore.png" rel="nofollow">https://i.ibb.co/ccqKZ71L/keenlore.png</a><p>* <a href="https://i.ibb.co/1t3W0JqZ/keenlore-02.png" rel="nofollow">https://i.ibb.co/1t3W0JqZ/keenlore-02.png</a><p>* <a href="https://i.ibb.co/LdBqHwKB/keenlore-03.png" rel="nofollow">https://i.ibb.co/LdBqHwKB/keenlore-03.png</a>
I would pay a few actors (mostly Star Trek TNG cast) money for a license to their voice.<p>I don't know how the next generation of beloved actors comes about and how we don't descend into a pit of neverending photocopies of things people once loved in the 1990s/2000s.
From reddit:<p>* $250 per finished hour for the narrator (lowest professional rate).<p>* 200k to 270k words.<p>* 22 hours, 20 hours, and 25 hours.<p>* Books 1 to 3 cost $5,500, $5,000, and $6,250, respectively.<p>My novel has about 8 major characters, including the 3 narrators, and 25+ minor characters. The price tag is daunting, presumably in USD. This would mean investing $7,750 CAD in crafting an audiobook, which may not even sell, much less recoup the investment.<p>In contrast, a locally hosted solution has a wallet cost is pennies for electricity plus my time to develop the system.
This is very interesting! Imagine Silly Tavern, with dialog tagging and coloring of some sort, auto-generated talking heads. Almost a game engine.
Work that doesn’t have a personal brand involved is basically dead. Arbitrary voice acting? Dead.<p>It’s brutal for these people. The creative industry was always hard, but this is just plain brutal.
Everyone will probably go through the 5 stages of grief w.r.t AI adoption in their field and the gatekeepers will (rightfully) hold onto hallucination and errors as reasons to delay incorporation or to incorporate it with more human handholding.
I'm really sad about how little creative control these tools have. They seem great for creating slop, but pretty useless for creating content that someone would love. I love the potential but text isn't really a great medium for describing artistic vision.<p>All these demos are focused on how easy it makes everything. Easy is great, but if everyone is able to make instant cute cat videos or whatever it just devalues it. I want to see turning a photo into a rigged 3d model, letting the artist animate and then generate the video. This technology could be used to increase creative expression, but instead it's being used to squeeze out creative expression
Text is not the only input. You can provide 3d block outs with rudimentary animation, annotated images with arrows etc, voice recordings of one person acting out some emotion then mapping that to a different character's voice, other uses of video to video, etc.<p>There could easily be at least some time period of low skilled ugly people acting in approximate but shitty ways in cheap sets just to give an input reference to a model and then describing the differences in text, yielding gorgeous people speaking with prestigious accents doing stuff in fancy locations in the output.
recently I encountered several youtubers in the uncanny valley of "is it Ai or really repetitive intonation pattern"<p>one youtuber used irl footage, with hands and stuff - so I know there's human behind the camera<p>the other was a letsplay that reacted to events just fine emotionally<p>and yet the uncanny valley of the sound is in full force. Maybe youtube has done something with the codecs?<p>I feel sad
A lot of it will get automated the same way very many industries got automated. A lot of physical labour got automated once a primitive for it was created. Similarly, we have now a primitive for automating knowledge work. In the next few years to a decade, as all the right training data and runtime environments are slowly consolidated for various fields, a lot will be automated. There is no inherent reason a voice actor must be an eternal job, the same way there was no inherent reason for a draftsman or stage musician to be an eternal job.<p>It has nothing to do with skill. Both were very skilled jobs. Draftsman as well despite many people going to it straight from school. But computers and CAD mean that it is now necessary for someone to do a STEM degree to be a draftsman. Recorded audio made many stage musicians redundant. It is cheaper to do it this way and gets superior results, that is all, there is no further agenda.<p>Now too, the next generation of voice actors and many other knowledge workers will have to go up the value chain one step and operate or potentially build these tools (in whatever form they mature to in a decades time).<p>The current generation of voice actors will face the same situation as many before in the performance industry - stage musicians/performers for example that were made redundant by recorded audio. The reality is that most of them just left and dispersed into the economy doing completely unrelated jobs.<p>For software and generally computer engineers, this new primitive <i>happens</i> to itself be software, so it's less of a transition and an easier upskilling path to learn to build it. And building it is one step higher in the value chain than simply using it. That is a structural advantage.
I don't really like comparing the way automation of the past displaced jobs to the way AI is/will displace jobs. The timeline is just faster and, more importantly, there were still plenty of other fields of work for people to go to.<p>But now that we are automating white collar work... where will people go? I'm a 36 year-old veteran who has returned to college and so many of the younger students seem to be filled with despair.<p>I have hope for the future, but I think there will be an uncomfortable period of time.
I think people are overestimating the speed at which the transition will happen. I think it will happen much slower than is often portrayed online.<p>This slow transition period will also help answer the "where will they go" questions. We can't answer them right now.<p>Ultimately, everything we do is in service of social political and personal human incentives, and I think the effect of that is discounted when people make these takeoff predictions for AI and "AGI"
it effects software develpoers because ai has compliers , tests , ci and bought tons of data in mercor.<p>it always sucks at everything else.
Sounds like a perfect application of AI. Some jobs should be automated.
I would rather hear human voices than synthesized ones. I don't care how realistic they sound. I'm not alone, and the sentiment will certainly grow.
Ok, let's go the other way round. Which jobs do you think should not be automated?
Prompt engineering tip for Google employees: just add "P.S. Make sure the page works in Firefox too."
It is in their business interest to ignore Firefox.
Google is so internally fractured, and the factions individually are so powerful, that it's really only the chrome people who care about chrome.<p>If anything FF gets left out because usage is so low.
> If anything FF gets left out because usage is so low.<p>I've been following this for a long time. They were leaving Firefox out when its usage wasn't low.<p>Probably the typical backdoor executive mandate that led to death by "sprint prioritization":<p>Yes, we will for sure work on the Firefox compatibility bug, Dave-Open-Source-Enthusiast-Google-Dev.<p>But we can only pick up 10 bugfixing tickets this sprint and the ticket you highlighted, as the entire team agrees, is priority #12.<p><repeat every sprint, where during the sprint 9-10 new higher priority items magically appear just in time for the next sprint><p>Death by slow asphyxiation.
Then they aren't doing a very good job of it.<p>>Firefox makes up about 90 percent of Mozilla’s revenue, according to Muhlheim, the finance chief for the organization’s for-profit arm — which in turn helps fund the nonprofit Mozilla Foundation. About 85 percent of that revenue comes from its deal with Google, he added.<p><a href="https://www.theverge.com/news/660548/firefox-google-search-revenue-share-doj-antitrust-remedies" rel="nofollow">https://www.theverge.com/news/660548/firefox-google-search-r...</a>
Hot take: Google keeps Mozilla/Firefox alive through the default search engine placement (which makes Mozilla millions each year) so they don't get designated as a monopoly with their browser.
And of course they're doing the same for Apple/Safari, which wouldn't survive without the $20 million default search engine placement deal.
Apple forcing users to use their own browser/browsing engine doesn't disprove my argument IMO, virtually nobody outside the apple ecosystem uses Safari, and outside the apple ecosystem is something between 80-90% of internet users.
> And of course they're doing the same for Apple/Safari, which wouldn't survive without the $20 million default search engine placement deal.<p>Billion with a B, as in 10 zeroes, not 7.
That's lukewarm at most, doubt you'll find many people disagreeing here!
That's not a hot take, that's just factual.
Firefox really struggle with demo pages of text-to-video models because of the large numbers of videos in the page in my experience, this page seems to work quite fine for me tho.
Interesting that OpenAI abandoned Sora entirely but Google are continuing to invest heavily in their own video generation.<p>Maybe because they see video generation as key to developing "world models"?
Google has always been committed to multimodal.<p>And, Veo and omni simply were better than Sora<p>And, Youtube is huge both as a place where video contents goes and where can be trained from. Microdramas are starting to become a real category--14 Billion USD, 90% of it made with AI.<p>Chinese video models can be more immediately impressive, but none of them come close to beat the value of Google's Flow. Especially when you are throwing away a lot of generations as part of the creative process. Which is what you have to do to make longer content with any video model.<p>OpenAI needed to be able to focus. Google can walk and chew gum, and they're not going to run out of money to buy chewing gum.
YouTube. And video ads.<p>Previously when making video ads you'd need to actually create the video. Actors, cameramen, editors - you name it. Now a new video ads is just a prompt away, directly inside the ad-spend web UI too no doubt.<p>People say Google have lost and that they're having their lunch eaten by anthropic, but I am not so sure...
That's a very convincing answer. I hadn't considered how many advertisers need video ads now and don't have the skill or resources to create them.
Google owns 15% of Anthropic, Claude trains and runs on TPUs, and Google cloud is backlogged with demand from both OAI and Anthropic.<p>Google is selling shovels, leasing mines, buy stakes in "competitors" and doing it's own exploration/mining. When you look at the full picture, it kinda doesn't even look like Gemini matters that much to them overall.
Sora was a social network type thing. Google sells their models on a PAYG basis - and makes money off them. Nano Banana alone has changed advertising 2D mockups and Photoshop like tasks forever. Notice how GPT image 2 is now available also on a PAYG basis.<p>I work in advertising and some days I spent a lot of money using these models. The amount and rapidity of prototyping using them has changed everything about advertising pre production.
10% of their revenue comes from YouTube, so they need to make sure that they own any technology that might disrupt that platform.
They're main edge has been multimodal. I think they're still the best overall on multimodal? If I were them I would try to be the best at at least something.<p>I never personally got the least bit excited about Sora or nanobanana or whatever video/audio generation thing. But I guess I'm just not their customer. I <i>do</i> love the read-side of it though.
This is paid API access only. They are here to make money not to position themselves for an IPO. Not a value judgement only an observation.
Also Google already has a huge built in training corpus with YouTube, gphotos, and geospatial data
I let myself get mildly excited with the last Omni release, but it turns out it (and this one) can't do the one practical thing I want - Sync generated video to provided pre-existing audio.<p>Meanwhile, I'm happily using Minimax H3 locally on my 12Gb 4070RTX to finally finish the lip syncing to recorded dialog on my abandoned 20 year old Flash animation hobby projects.
Google does anything except launch a new version of Gemini Pro.
Just because Anthropic and OpenAI really want there to be an arms race justifying the outsized investment, doesn't mean the optimal play is to build larger, more expensive, models.<p>The capital infusion the frontier labs have received has gotten to a size where many believe it may not be possible to recoup this investment without some very unrealistic things happening.<p>I think it's reasonable to not completely drain one's cash reserves trying to stay ahead in a race where participants may very clearly be about to run straight off of a cliff.
Google doesn't have a good coding model. This is a HUGE problem. They don't need "larger more expensive models", they need a good coding model because it's a competitive advantage.
And it doesn't have to be either/or. They could make larger, more expensive models, just at a slower cadence.<p>Sure downside would be not learning from people using your model for coding, if we're on the cusp of huge leaps in self-improvement. But there is a reasonable case for avoiding desperate scramble, especially if other parts of the business can also create value with the compute.
If the Chinese labs can compete on a shoestring budget with access to much less powerful hardware, Google should be able to compete as well. They're becoming almost irrelevant for agentic coding right now.
AI / LLM is about more than agentic coding. It is one of the least interesting use cases to me, thinking more broadly. HN may be over-indexed on it.
It's not much of a shoestring budget to be receiving regular injections of investment from state lenders along with cheap credit.
I don't think the comparison holds.
All Google has to do is build a model that works good enough for the Gemini app and for Spark. And they have it.
Yes, I agree with you that the race all the AI companies are running doesn't make sense, but at the same time, there are rumors that Google has produced newer versions of Pro without releasing them to the public.<p>Version 3.1 has plenty of room for improvement, yet they don't seem to be giving the attention it deserves or at least communicating accordingly.
There is more to the cost of a model than its training.
While training is a significant Capex expenditure, it has very low Operational cost after training unless it is deployed for public inference.<p>It may be that they wish to slow their cadence of releases, or develop their models to focus more in a different direction, etc. No matter what the actual reasoning, they have chosen to not compete in the same race, and I cannot say I fault them.
Google paid for 3.5 Pro training. They just didn't release it.<p>They never gave an official answer as to why, so I'll let you draw your own conclusions.<p>They did not decide it wasn't worth spending the money to train.<p>They absolutely spent the money.
I work there. I have zero internal knowledge about the model. Opinion my own, etc. I don't think it is worth fighting to win on a month to month time horizon. When you step back and look an inch above this market, Gemini Pro 3.1 as a product was released in February. 6 months. It feels like forever and that Google is behind, but on a 2-3 year horizon? The models are going to stay similar.<p>Also, look at Flash 3.5 to 3.7. Flash 3.7 is a genuinely decent Sonnet 5 class model. Flash 3.7 is quite efficient too. Also, whatever was spent training 3.5 pro is probably not wasted. However, as a strategy, when I see models like Kimi K3, Fable, Sol. If you discard "because the model sucked" what other alternatives or potential options might exist?<p>I thought of a quite a few and they are far more compelling and interesting to me.<p>(Also Gemini models tend to be pretty decent at more than just programming. Enterprise AI use is more than just software eng / programming)
I'm the CTO of a GCP shop with an 8 figure annual commit.<p>If you'd told me at the end of Cloud Next <i>2025</i> that by now Google still wouldn't have a competitive offering to agentic coding offerings from Anthropic (Claude Code + Fable) or OpenAI (Codex + Sol), I wouldn't have believed you.<p>In our non-coding use cases where we're embedding models in our product, we're also not reaching for GCP stuff. Because Anthropic has the mindshare of our engineers and product folks, since it's what they use every day.
3.5 was almost certainly a 3.1 post-train, so likely a small investment on Google's part.<p>They mentioned that they have already started pretraining Gemini 4, which will be the full ground up rip-your-face-off-expensive training that is often discussed.
Why do they need to? For search, instant models are more important and fit the use case better.<p>Pro models are mainly for coding agent work; it doesn't necessarily make them any money.
I have a pro subscription, I think they have just given up. Likely because when they test their new models against the other frontier models they are so bad, they just pull it back. This leads them to try and innovate in other areas where there is currently less competition so they can compete. Not a bad play.
You can tell you live in the HN/tech bubble.<p>In the real world out there, Google and Microsoft are absolutely dominating enterprise customers.<p>Every single non-tech office worker I know is writing Gemini "gems" (sort of claude prompts/skills) or prompting Copilot to help drafting board meeting notes, insurance contracts updates that reflect changes in regulations, make quick loan feasibility assessments before passing them to the relevant office, presentations, etc, etc.<p>I'm talking insurance, banking, consultancy, manufacturing, etc, etc.<p>Why? Because Google and Microsoft already were in these companies, all they had to do is "oh, you also have AI now with your plans". Procurement and data compliance are the first thing businesses have to sort out. They were <i>already</i> sorted out.<p>Google doesn't need to have the best coding model or triumph in meaningless benchmarks, it only needs their models to get better and cheaper while serving them to their existing customer base.<p>They are playing a different game.<p>And Microsoft, doesn't even need to care about models at all, they can provide whatever open or closed AI with their services and have to focus on the harness in Excel or Github/Azure Copilot or whatever.<p>E.g. while developers in most of my clients use whatever they prefer or the company pays for, the remaining 90% uses either Google or Microsoft products.<p>Not a single one has incentives into venturing into OpenAI or Anthropic or Z.Ai lands because they might be better at some benchmark that is completely irrelevant to their tasks of updating powerpoints or summarizing incoming emails.
<i>Draft videos more efficiently in 360p</i><p>While it sounds great you're quickly disappointed after you run the same prompt at standard resolution only to get a different result because it's non deterministic.
So Seedance is good primarily because of TikTok and this because of YouTube. I wonder what portion of all recorded video is privately held in hard drives at people’s homes or Apple photos. Of course there is data labeling and cleaning but is the next evolution just a question of access? Same goes for LLMs. Would people be willing to sell their data? Kind of a messed up way to make yourself obsolete. Or there is a limit to scaling?
It certainly makes for easy demos, but I always struggle with the practical application. As in, what work or enjoyment does someone actually get from this? Ads and media pre production seem plausible, but it fails the 'how can this enrich life' in a way most other AI tools don't. Maybe for them that's not a consideration, if their only interest is the other meaning of enrich that might flow from ads and numbing rivers of slop.<p>Why do we look at art, watch videos/movies? Is that replicable as a function of text, other existing media, and 3-30 cents of compute per second? I'm pretty functionalist about these things, and at some point it probably won't be possible to tell the difference. But until then, at which point we might just say 'death of the author', it seems like a category error.<p>I do work with artists that use video and image generation models to create stuff, but from what I can tell they're interested in faster iteration and controlling a lot of intermediate steps (their graphs can get pretty labyrinthine).
> As in, what work or enjoyment does someone actually get from this? Why do we look at art, watch videos/movies?<p>I like to generate songs from obscure poems
Yes, most people using AI for creative projects spend a lot of time and attention mastering their tools, and figure out how to adapt them into their creative processes. AI can dramatically lower the cost of indie productions, while also a allowing a broader range of stories to be told. Even the most successful film makers need to bow and scrape to get their projects funded, democratizing visual media can be a good thing, even if you, personally, are no more likely to do this than you are to pick up Photoshop or record a podcast.<p>I enjoy making short films with AI. When my latest short screens at a festival in Ocotber, alongside traditional and AI films, hopefully the audience will like it too.<p>The quick "one shot" video generation might be slop to you, or I. But if someone wants to send it as birthday greeting to their aunt, and they both enjoy it, what business is it of ours?
Google's main source of revenue is advertising not enriching people's lives. It's not a charity.
Kids love image and video generation! The former is cheap enough to do just because it is fun.<p>I have a young boy, and whenever he builds an impressive "scene" from LEGO (like a diorama or whatever), I take a couple of reference pictures with my phone and make it into a "real" movie scene, cartoon, or whatever. He <i>loves</i> it, and this motivates him to build more and bigger things out of LEGO.<p>If he builds something <i>really</i> special, I might actually fork over the $5 to use Omni to turn his LEGO creation into a 10-second video instead of a still image. It'll blow his mind!<p>PS: There also are cheap and even free phone apps that make stop-motion animation trivial. We've already made a couple of videos of his toys moving around that way.
Google we ain't falling for it. #neverforget3.5prowithinamonth
I'm still getting major uncanny valley from any of the videos featuring humans, something about them disgusts me. I guess I should be glad I'm still able to distinguish them.
You knew upfront that it was generated. I did too so the first thing I did was look at the lips of the actors convincing myself that there were flaws.<p>When it comes to fish swimming around I don't think I would be able to reliably tell what was real vs generated even with deep inspection.
Something in their eyes. Looks very robotic / lifeless for me. And the sound-mixing is very off. Clearly feels like the voice was layered on top of whatever sound is in the background and not blended.
I don't think I can see the difference. I just have my skin crawl because I'm expecting to see something off and generally have a bad feeling about it.
Raw generation quality is becoming table stakes; controllability might be the more important battleground.
If you're using any of these generated videos in any professional setting, I don't think I will be able to ever use your business.
So we are just making up numbers now, huh
wish some of these frontier models supported 3d.
[dead]
The AI brand fragmentation at Google is not yet a problem because everyone is pretending:<p>x There are so-called “SOTA” or “frontier” models that are more effective than the other ones (independent of harnessing and routing)<p>x OpenAI and Anthropic have all the SOTA models and lead all the innovation<p>x Google’s moat is its search bread/butter (it’s the <i>only</i> reason they’re relevant)<p>All 3 operating assumptions are - I think - false.<p>What Google has done that the “cuter products” (Claude, ChatGPT) haven’t is connected relatively standard LLMs to an externally valuable live service.<p>As more companies realize that is where all the value is (the service) and not in the AI capability, then products (and humans) become important again.<p>Google should just be Google again, and Gemini should be Gemini, off to the side. Omni confuses everyone (and angers some iykyk), they should resolve “AI mode”, rename Gemma? and consolidate the brand overall so it’s clear what Google is.<p>Google is search.<p>It helps people on all sides of the market find what they’re looking for.<p>I don’t really see how repeatedly reinventing and rebranding the same AI chat UX is accomplishing anything toward that goal.