Hacker News

Top stories

Live mirror
30 storiesupdated just nowView source snapshot
  1. F-Droid 2.0 (f-droid.org)
    288comments
  2. Show HN: Make cursed fonts like Times New Bastard (mitpit.com)
    80comments
  3. Goodbye Google (ocallahan.org)
    3comments
  4. Show HN: Whiteboard (YC W26) – An open-source IDE for thoughtful software design (github.com/devdotfast)
    95comments
  5. Why is the liver so weirdly regenerative? (dynomight.substack.com)
    176comments
  6. 2DWillNeverDie (2dwillneverdie.com)
    23comments
  7. Fearless SIMD v1.0 (linebender.org)
    34comments
  8. Rails World 2026 Opening Keynote [video] (youtube.com)
    311comments
  9. Toyota is taking the Corolla electric (electrek.co)
    465comments
  10. My weird new hobby: Wandering around Tokyo on Google Maps (ahmedhossamdev.com)
    117comments
  11. Using LLMs to trace alchemical knowledge and decode 17th century letters (resobscura.substack.com)
    19comments
  12. Google’s Project Suncatcher to put ML infrastructure in space (blog.google)
    281comments
  13. Two-tier encryption in the UK (macanorak.com)
    384comments
  14. Writing Parquet files using Haskell (datahaskell.org)
    3comments
  15. Book review: Is parallel programming hard, and, if so, what can you do about it? (ahelwer.ca)
    33comments
  16. The Board Game of the Alpha Nerds (2014) (grantland.com)
    26comments
  17. California is chasing wealth that has feet (landeconomics.org)
    556comments
  18. Show HN: Air-gapped file encryption as self-decrypting HTML page (apeleg.com)
    17comments
  19. Security auditing in the age of (good enough) AI (trailofbits.com)
    10comments
  20. Sourcehut account takeover via build logs (XSS in ansi2html) (blog.arusekk.pl)
    17comments
  21. Show HN: Koi.rest – watch some fish and regain your balance (koi.rest)
    41comments
  22. The forgotten battle of East Lansing (eastlansinginfo.news)
    15comments
  23. WaveDigger: Dig into wireless signals to discover their physical locations (github.com/christianrowlands)
    20comments
  24. Opus 5.5 is good at explainer videos (launchvideo.io)
    108comments
  25. Forging 1024-bit RSA signatures in nearly SNFS time [pdf] (iacr.org)
    9comments
  26. Stable (YC W20) Is Hiring Product Engineers (usestable.com)
    —discuss
  27. The Bayeux Tapestry: Woven by the Victors (historytoday.com)
    3comments
  28. Geothermal heat map of US hot springs (soakingsprings.com)
    39comments
  29. Nokia Design Archive (2025) (aalto.fi)
    120comments
  30. Motor Characterization for Small Running Robots (2016) (robot-daycare.com)
    2comments

Opus 5.5 is good at explainer videos

189 pointsby 9h agolaunchvideo.io
108 comments
8h agoHN ↗

What is the agent using to make those videos?

8h agoHN ↗

Wow, this is phenomenal. Thanks for sharing.

If you're the author congrats, great work!

8h agoHN ↗

Arguably, this type of content was already slop before the AI times, so nothing of value was lost.

8h agoHN ↗

Right. My first question when I see this type of video is asking for a text write up. What tool is used to create these videos doesn’t change that.

8h agoHN ↗

Can I just say "oh dear god no ..."

(not that they were great before, but ...)

8h agoHN ↗

Like many of us, I've been hobbying away on some SaaS that wraps an LLM. But I often worry, isn't this just adding buttons to an LLM call? What's the real value here? Quick money grab and wait for copies? Or do my 20 years of saas and business development actually make my SaaS better then others?

8h agoHN ↗

Even if I could answer your question (which I can't), things change so fast right now that it might not be valid for long. All you can do is be willing to rapidly adapt and stay focused on where your product adds value.

8h agoHN ↗

I believe it completely depends on your values and perspective, as a poor analogy: when mobile apps were a novelty a bunch of them were just that, novelties, the beer drinking app, the flicking a Zippo-esque app, and many more that hadn't much functionality, were more a tech demo, cool toys, etc.

If you don't care so much about adding something significant to the world, and your LLM-buttons-wrapper is a novelty, chug along and see where it goes if you're having fun. Looking for what value mean at this moment is a much harder effort, and probably a luck-based endeavour.

I don't have fun creating LLM-wrappers or at least haven't thought of a fun idea for it that could be even a fun novelty to work on so personally I'm not doing it but I'm using LLMs for other fun stuff that required much more of my free time before.

7h agoHN ↗

I used to think like this, but eventually I realized it's mostly just a way to pat yourself on the back. And ironically, I rarely saw people doing significant things waste energy talking like that, since there's no benefit in the exercise.

Very few things can't be reduced to insignificance. MacOS was just cribbing PARC, Facebook is just a glorified PHP forum, Dropbox is just SFTP, etc. etc.

Even in deep-tech, Zipline is just wrapping from deeper-tech (batteries motors etc), GLP-1s were a VA throwaway that dusted off, etc. etc.

LLM wrapper doesn't mean anything, it's an implementation detail. Besides being an LLM wrapper what is a given thing?

The beer app wasn't just a beer app, it was the intersection of the first time accelerometers were doing something that the average consumer could interact with in their pocket, the first time there was something to spend money on for your phone besides a wallpaper, a ton of things.

If anything now when something seems trivial or like a toy, but it has a lot of traction, I want to spend energy figuring out why is it more meaningful than it appears.

8h agoHN ↗

Many of us have the exact same thought process. Related: I have heard indie game developers have stopped the traditional practice of updating a dev log as a form of marketing and community building because people will point their LLm at it and create your own game before you do

8h agoHN ↗

I have some strong feelings about this particular story (people point LLMs at your devlog)…

- Video devlogs were long a shitty way to make a community (I can explain this one in more detail, but in short, the tradeoff between opportunity cost and video quality is a Pareto curve that is bad at all points on the curve, if your goals include both “make game” and “build community for players with devlogs”, the way out is to drop one of those two requirements, either “make YouTube/Twitch/TikTok channel” replaces “make game”, or you build some other audience for your devlogs)

- Good devlogs don’t have that level of detail, to let people easily recreate games

- Most games can’t easily be replicated by LLMs, in short, the people who are good at steering LLMs like that are making their own games.

This is not the first time I heard this story. Sometimes it comes with the lamentation “back in the day people could build communities with devlogs” and no, that was generally not a good way to build communities. It was mostly good streams from YouTubers who were making the game in order to make YouTube content or it was mediocre streams from random devs. Throw in a few people who are famous game devs who choose to stream and get a big audience because they already have a community.

8h agoHN ↗

Good devlogs don’t have that level of detail, to let people easily recreate games

The whole problem with LLMs is that you don't need the detail.

7h agoHN ↗

Ok, so let’s get some data points. Are there games being recreated from devlogs? I’d like to see what the recreations look like.

7h agoHN ↗

Ok, so let’s get some data points. Are there games being recreated from devlogs? I’d like to see what the recreations look like.

An LLM "recreation" doesn't have to be complete or very good. It just has to steal enough thunder to be profitable.

7h agoHN ↗

From what I’ve seen, most of the flood of LLM-created games aren’t profitable, just like most of the games people make in devlogs. There’s very little thunder to be stolen in the first place.

7h agoHN ↗

Details have a huge impact on how fun a game is. So far the vibeslop games I've seen look about as fun as crappy unity assets flips or mobile games. every so often the Instagram algo will show me some post a long the lines of "The game industry is finished!" and there is no way anyone can seriously looks at the games in those posts and say they look fun to play.

And someone trying to vibe copy another person's game is definitely the type who has no idea of value to add. So many of the slop games the whole concepts are so generic that they definitely just asked chatgpt for everything.

In the modding scenes for some games I play I've seen vibe coded mods where the gameplay additions make no sense and have no sense of balance or fun, with these completely new to the community devs having ko-fi links set up from the start.

7h agoHN ↗

What I think of is all the HD remasters out there which look disastrously worse than the original, at least in my eyes. Why does this happen? Are the best artists working on new games instead? Are remasters pushed out with less care, shorter schedules, and less budget? Maybe some combination… and maybe there are some parallels with quick, mostly unsupervised LLM copies of a game.

6h agoHN ↗

If someone makes a cheap copy of my unreleased game and it's bad, that's also bad for me, I think.

5h agoHN ↗

I think the much bigger problem is that these people are usually trying to do some "passive income" play, drop-shipping type crap. And will soon be spamming what would be the usual discovery mechanisms with slop. It's happening on youtube for educational type videos.

7h agoHN ↗

It doesn't need to recreate your game. It can still muddy your marketing and redden the ocean. Especially if someone is in the business of trying to front run an indie game, they don't need to be any where close to a full or good game to attempt to steal any mindshare your project had.

6h agoHN ↗

Is this happening at any significant rate?

AFAICT this is a fear that some game developers have, that someone will steal their game and run with it, and the stories told are more repeated based on this shared fear than based on realistic scenarios. People repeat it because it’s a good story.

And yes, I’m aware of some situations like 0x10c and the like.

5h agoHN ↗

I frequent a couple of game development communities, and it's kind of hilarious that some people won't even talk about their general game concept for fear that others will steal their idea. Like wow, you're making another rogue deck-builder, I have to steal that!

The idea is not valuable, ideas are a dime a dozen. The actual value is in your implementation.

4h agoHN ↗

Is that true anymore, LLM will implement everything for you, idea actually matters more now

1h agoHN ↗

LLMs will generate. Taste continues to dictate what to keep/discard. (Note that taste can get you to good, but good won’t get you to money on its own)

11m agoHN ↗

This is fantasy and if it was true the idea would be public when the product was released.

There's no moat in an idea either so I guess it's down to marketing budget.

5h agoHN ↗

The appstores are full of knockoffs and even resubmitted decompiles. I certainly wouldn't put it past the scammers.

More specifically to AI... A blog on a nice website is no longer a signal for quality so there's just no point in devs investing in that.

7h agoHN ↗

I've been vibe coding a game on the side.

It really makes me appreciate how important good game design is. Claude is doing a fine job coding everything I describe, but it doesn't really understand fun, so I need to.

(It's not a great game)

8h agoHN ↗

indie game developers have stopped the traditional practice of updating a dev log as a form of marketing and community building because people will point their LLm at it and create your own game before you do

And that's why we can't have good things...

(This is just gonna keep happening more and more until eventually we'll need something like a patent system for ideas)

8h agoHN ↗

You could build something really useful, but if I can replicate 95% of that in a weekend with Claude I'm obviously not going to pay you for it.

Some SaaS products have network effects, but that doesn't appy to new software.

Starting a SaaS business in 2026 seems a bit silly to me. Not saying people won't make money, but you'd have to be pretty lucky.

7h agoHN ↗

You'd be surprised how much money companies throw at very simple products, I'm working for a company making 100k mrr and was virtually entirely vibe coded by one dev until very recently

8h agoHN ↗

It's not always a call that wraps an llm.

The LLM only does average returns from average inputs.

8h agoHN ↗

I don't think "is my SaaS better than others" is the important question, instead "is SaaS going to be valuable in a cheap bespoke software future".

I've been playing with an LLM backed reference checker for my partner who is a lecturer. For each reference in some work she's marking an agent is dispatched to read the referenced source and check that the reference is correct / accurate / not hallucinated.

It's useful to her, but it feels like it would be a waste of effort to turn into a product, because pasting the same document into Claude/ChatGPT with a prompt like "download and read each reference etc" would work pretty much just as well.

I had similar concerns about an AI backed training generation company I interviewed with. I asked how they saw themselves competing with increasingly capable generic harnesses. I don't recall exactly what they said but the impression I got was they were so focused on competing with the big LMS players that they hadn't really considered it.

8h agoHN ↗

this is why people have been suggesting an AI bubble, outside of the frontier model labs

8h agoHN ↗

Typically maintenance, reliability, support and liability; nothing new in SaaS land. Lots of companies have the money to hand-roll things that are provided by some external SaaS but would much rather it was someone else’s problem.

Also for creatives, some sort of aesthetic control? Most of the examples on this page look like templates you’d see in an office suite that appear swish, but are ultimately bland.

7h agoHN ↗

I've been struggling with that.

I built https://gallotails.com/ because I am a cocktail enthusiast. First commit, March 6th 2019. One of the features I am proud is which cocktails you can build based on your ingredients, and I was smart, the bar ignores garnishes and can suggest cocktails you're one ingredient away. But you have to add the ingredients. And interpret the bar screen.

And here's what I did with ChatGPT a few days ago: https://chatgpt.com/share/6ab59877-4b74-83e8-8da5-966d9f3a38...

Clicked the microphone icon and rambled for a few seconds (could have just taken a picture too I guess). Nudged ChatGPT to only give me simple cocktails I can stir instead of using a mixer. Said I could buy a lime, sure. Then asked how to keep the ginger beer fresh.

There is NO WAY I can code all of that in my little gallotails.com - a single chat window has entirely replaced my site. And that's my pet little project I do it on weekends, mostly to keep my tech skills sharp. I can only imagine what the biggest cocktails website, actual companies, are going / will go through once more and more people just realize the chat tab is enough. And I am not talking about purely content, ChatGPT can do everything my site does, and more, into any direction, in an instant.

Which left me with the question, then what should my site be? Community? To be taken over by agents?

Currently munching on what sites like mine should do, and what they should be.

7h agoHN ↗

Then maybe the need for your site is gone. That’s a good thing in the sense that the problem is solved. If your goal is “do something in the cocktail space” (which is also roughly the companies’ goal I would say), indeed you should do what the AI can’t, which I’d say is be human touch. An AI can’t give your opinion or view.

7h agoHN ↗

Then maybe the need for your site is gone.

Oh yeah, and I am not mourning or anything like that. And I do want the human touch, because AI can't. But humans, at least over the internet, can send their agents, or companies can run agents and I won't know from my site if it's an actual human, or some AI fabricated interaction (can't even trust an image it uploads).

When AI first showed up a few years ago, I thought about opening an actual physical cocktail bar. Hypothesizing humans will be tired after a long day of their agents talking to other agents, they would crave for actual humans. I would ban electronics inside - no distractions. Maybe one day I will go for that...

7h agoHN ↗

I personally am more interested about personal opinions about things like cocktails rather than whatever statistically average cocktail the matrices of chatgpt come up with.

Chatgpt would shit on my great grandma's waffle recipe but it's the one I'll keep making till the day I die.

7h agoHN ↗

I agree there will always be a need for community. Human connections. Folks experiencing the tastes, smells, etc.

Eg from your share, the model output “I prefer it Scotch-heavy rather than 50/50 because otherwise it gets very sweet.”

It’s basically cribbed this from someone; it literally can’t taste.

The frontier labs are not yet able to do rlvr over mixology/meatspace. So maybe think on how to capitalize on that.

22m agoHN ↗

Books, writing, comments, etc on cocktails are certainly in the training set. The model doesn’t need a sensation of taste in order to understand what humans enjoy. RLVR specifically on cocktails is not needed as the general intelligence of the models increases

7h agoHN ↗

Here’s an easy proxy: Did you feel the need to build it because nothing to your liking existed? Did it take some effort to build? Do others feel the need to use your thing for those reasons?

Then it’s sufficiently non-obvious. I believe there’s still a pretty big field of things that fall into this category. Making proper things is a lot more than writing some code.

7h agoHN ↗

Or do my 20 years of saas and business development actually make my SaaS better then others?

Nope. If all your doing is prompting an agent to build stuff (especially if what it is building is based on LLMs), then you might as well take it back to the whiteboard and rethink how you're spending your time/tokens.

7h agoHN ↗

Like many of us, I've been hobbying away on some SaaS that wraps an LLM. But I often worry, isn't this just adding buttons to an LLM call?

That's why I created isitsaas.io: your platform for determining whether your llm wrapper has product potential, or whether people would rather just use claude directly instead. Simply pop in your business idea, and our award winning, proprietary technology will ask claude if you can make money off of it.

Trial memberships start at $10/month

6h agoHN ↗

You are thinking from the point of view of an engineer, and risk falling into the rsync dropbox fallacy.

Start from thinking about product value add, and user behaviors first.

3h agoHN ↗

Making the edge cases work, and lowering the cost of running it via good model choice, context management etc is in some cases really hard. That's valuable in a decent number of cases. I have a system that does something in financial markets - making sure it doesn't screw up, and doesn't cost a fortune to run, is the entire thing for me.

8h agoHN ↗

I don't want to be a hater but I made this a while ago and I honestly think the videos are a good bit nicer than what you are creating with Opus 5.5. https://mnfst.video/

8h agoHN ↗

This is just taking an existing presentation slides skill and generating a video out of it. There are no drawings and animations that explain the actual concepts.

8h agoHN ↗

I wonder what the future is of SaaS in a world where everybody can vibe up what they need on demand.

8h agoHN ↗

Yeh try that. Not that simple at all, this is why servicenow and sales force are just killing it

7h agoHN ↗

Looking at the absolute dog shit Ai menus and posters I see everywhere I can guarantee you a good 70% of the population is physically unable to produce good things no matter the tools they have access to. They can't even be bothered to add "make it look good" to their prompt

8h agoHN ↗

Not sure if it is working for others, but if I point it at any of my domains I get "The model provider rejected the request".

8h agoHN ↗

These aren't even that good in comparison to the stuff being posted on Reddit and twitter/X.

Some of the stuff I've seen is mindblowingly impressive in how they've captured quality creative decisions, and how the models can reason about visually appealing designs.

It understands things like continuity, themes, facial expressions, cultural memes/references, symbolism, etc. And it generates things using a lot of multi-modal behavior by using 3d, images, etc.

Some really incredibly impressive stuff.

8h agoHN ↗

I won't open Twitter, but where on Reddit is this stuff posted?

5h agoHN ↗

Interesting. Bit odd that the head doesn't occlude the background (it's a circle, not a disk), and not sure about the title font.

The drawing/animation style feels suitably off from xkcd - more naive stick figure, than what xkcd looks like?

Some speed/timing issues with the walk cycles.

For all that - mind-blowing that it's generated so quickly.

51m agoHN ↗

Thanks. I’ll pass the feedback on to Opus 5.5 :-)

4h agoHN ↗

Dude, so well done.

Say more about the process and / or drop some link pointers.

58m agoHN ↗

There isn’t that much to say. Start with an empty directory, put the XKCD image in it, add your ElevenLabs API key, and set up a Python virtual environment. Install ffmpeg.

Then start Claude. From that point on, it’s mostly prompting and evaluating the results. Like any good engineer, you shouldn’t simply start with “Create a video from this comic.” Instead, take a step-by-step approach: discuss the visual style, animations, ask for audio examples, create a script, and then develop a storyboard (ask for a .html document).

The important part is to break the whole generation process down into small, manageable steps. I also used the brainstorming skill from Superpowers for the discussions. It makes the whole process much easier when the AI asks the questions and I just have to answer them.

The result is 4500 line python code generated by Opus 5.5 and a lot of video and audio files.

2h agoHN ↗

I had already seen the comic, but watching the video reminded me of "The Last Question" by Isaac Asimov [1]:

After all, I undertook to tell several trillion years of human history in the space of a short story and I leave it to you as to how well I succeeded. I also undertook another task, but I won't tell you what that was lest l spoil the story for you.

[1] https://users.ece.cmu.edu/~gamvrosi/thelastq.html

8h agoHN ↗

lol no it's not.

I asked it to make a site for my TI4 reference project: https://axiomvortix.com

It produced a video about an AI startup with the vague goal of centralizing data.

8h agoHN ↗

My friend..definition of “good” is very relative. These are just flashy slide decks

7h agoHN ↗

This one's been stuck in my head since yesterday. I just keep thinking about how long this woulda taken in After Effects. So I was experimenting with having Opus 5.5 make these kinda animations in javascript and realized I can get it to export a JSX script for AE. So I can have it take the first swing, then bring it into the tool I actually use.

5h agoHN ↗

Curious about the song - the mp3 in the repo is wildly different from the linked YouTube recording?

8h agoHN ↗

I've been experimenting with AI video generation beyond raw video gen for some time now. Opus 4.5 was the first tool I tried last year, and it was okay...

Fast forward a few months, and Opus 5.5 is now running my entire video production pipeline. This model is leaps and bounds ahead of the previous model.

I can generate a complete 5-minute explainer video on my MacBook Air M4 in about five minutes from a single prompt, using the Gemini TTS API and MiniMax H3 Max for b-roll. It creates thumbnails and titles, handles the entire process, and automatically uploads the result via the YouTube API—all powered by a Claude Code plugin. (youtube.com/@ctrlaltexplain for those interested).

7h agoHN ↗

Recent models have been a big step up on graphics generation. Not exactly sure why, it would be neat if they would release any quantity of technical blogs.

6h agoHN ↗

FWIW the examples in OP were generated by image models, not Claude - Claude just orchestrated the other models via OpenRouter

28m agoHN ↗

Emergent capabilities, bitter lesson, increasing general intelligence, etc

7h agoHN ↗

This is so different from the OP's site. This is so much better, trying it out now.

5h agoHN ↗

Did you get the raw js artifacts to continue working, or making "sequels" - or just the video file?

Does the js stuff render real-time?

Did you get script/dialog and voice settings (again, for making other films in same series/theme) - or just the rendered audio?

2h agoHN ↗

AI doing things that creative people do makes me sad. Stole everyone's art to regurgitate it on command to billionaires richer, and save businesses the expensive of hiring someone with actual talent to do it.

I don't care how "good" the graphics become - it will always just be slop. This isn't the part of our lives we should be trying to replace with technology.

8h agoHN ↗

Yeah, and other models as well. YouTube is now full of AI explainer slop videos.

7h agoHN ↗

Today's free videos are all used up. Deploy the agent to your own OpenComputer account and it runs without limits.

So predictable.

7h agoHN ↗

I had Sol 5.6 about a month ago populated a test account, take screen shots, write a script, used ElevenLabs TTS, with Remotion. It did an amazing job.

7h agoHN ↗

Building civ.game with Opus. Its fully procedural, threejs plus wasm. What a time to be alive.

7h agoHN ↗

I suppose we should've seen this coming (and many did) in the 80s. This is the late stage of digital and computerised media. We conned ourselves for a few decades into thinking this was a new medium for art or a new frontier of human communication, but in reality we were just training a simulacrum of reality, devoid of meaning or intent.

7h agoHN ↗

These explainer videos are garbage, whether done by humans or not. Just like those crappy Netflix "documentaries". Don't get me wrong, it's impressive, but it's just a more automated way of outputting low effort, low quality content

5h agoHN ↗

It is kinda funny I am always interested in Anthropic's new releases, but i have a great distaste for their actual release posts/videos. They LOVE these stupid style videos for any feature they release, their text posts are usually just a ton of nonsense that i dont wanna read, and now there is this huge "revelation" that their models can pump this annoying format out.

I say Bummer if more people adopt this style, its gonna become less and less authentic feeling as the months go on now.

7h agoHN ↗

I tried it, but I can't see the video. The model just outputs this in the chat: /root/workspace/renderer/out/local.mp4

No way to download or watch it in the webapp of opencomputer. No instructions on what to do? Do I need to run it locally? It seems it did use usage.

6h agoHN ↗

Ask Claude for instructions to play the video.

7h agoHN ↗

Indeed, it's really great. It does the smartest thing of all: it creates an HTML page and uses raw JS timed animations to make it work. Then it records it with ffmpeg and saves the output as the final video.

About 6 months ago I spent 12 hours creating a video for our Edge.js [1] announcement that, in the end... people were not very inspired by (to say the least). See the video here [2].

Then, about a month ago, I spent 6 hours creating a video for our Wasmer SDK announcement. It was better, but still didn't go as viral as I wanted [3]. I always thought we would need a big budget to do them.

But then, yesterday I tried Opus 5.5... and man, I'm impressed. The video was one-shotted with this prompt:

    Make a modern slick and punchy video for this announcement:
    [content of the blogpost in markdown]

This is what Claude Opus 5.5 one-shotted (tl;dr: we went viral): https://x.com/wasmerio/status/2102849543260029379

[1] https://edgejs.org/

[2] https://x.com/wasmerio/status/2033966082944577693

[3] https://x.com/wasmerio/status/2094845905379922302

7h agoHN ↗

>> records it with ffmpeg

Exactly how? What does ffmpeg record?

7h agoHN ↗

I believe it captures each of the frames as png and then ffmpeg puts them together as a video. Although someone from Anthropic may be able to explain this better

7h agoHN ↗

All of these already have a similar vibe and will become commonly recognised in a few weeks time, making them not good.

5h agoHN ↗

Kind of clever. Not sure what this has to do with Opus 5.5 though. Creating an html presentation with transition animations has been a thing since Opus 4.6 - Recording it with playwright/ffmpeg doesn't seem like a model thing.

5h agoHN ↗

I'm looking to put together a tutorial discussing fundamentals of Blues and Swing partner dancing. I have the raw source material. I'd like to change the background, change the attire of instructors (myself and a friend), and fix up other issues (dead space, umms, and other gaffs). The footage is about 22 minutes, and can be split into 5 different segments of about 3 - 4 minutes each.

I'm running Arch Linux and would rather not install DaVinci Resolve, but Blender would be fine for the non-linear video editor.

My plan is to use ChatCut to trim and split into the different sections. After I have the five segments, what would you suggest for swapping the background (environment) and attire? There are two simultaneous camera shots (front and side) that I'd like to stitch together as well.

Any suggestions? (Paying someone a couple of hundred $CAD to take this task off my plate would also work.)

3h agoHN ↗

I have an OSS framework for this: https://github.com/scosman/videowright

- Voiceovers: aligns animations to the voiceover, can generate voiceover with elevenlabs, or will transcribe and timestamp a real voiceover

- can reorder scenes both in code, and using ffmpeg for audio.

- interactive controls during authoring, can ask for micro edits or re-builds

- MP4 export/encoder

- Generates the video with agent of your choice (obviously)

3h agoHN ↗

Lot of these demos tend to squeeze in as many animations and transitions. Its usually hard to take away anything from the video at the end.

I wish more people optimized for learning than "how fancy can i make it look"

2h agoHN ↗

Ah remember the good old days when models couldn't do images with text because it couldn't spell or letter would be borked?

2h agoHN ↗

for some reason, claude models are extremely bad at manim, including opus 5.5 (surprisingly gemini is very good at it)

1h agoHN ↗

I tried Gemini (3.1 Pro) for creative writing a couple days ago and was absolutely blown away. I have not seen one this good since ChatGPT's initial launch day, before the rounds of lobotomization and RLHF.

I am not super sure what makes it so good, but it seems like it's a good option to try if you have access and want to see how the frontier models are doing.

Maybe this is why Apple chose that family to help train Siri AI. (Siri AI is not a fine-tuned Gemini)

1h agoHN ↗

I work for a company doing marketing (I make ads). Essentially I've trained my OpenClaw machine to know everything it needs about the company, know what will perform best, and then it writes a script, creates an avatar, generates videos, b-roll, edits them together with captions and CTA end card. And they've performed better than when I was doing it all manually. Pretty crazy.

1h agoHN ↗

Not a comment on this link in particular, but just a general observation from my experience.

One should be skeptical of the demos they initially see at model release, as there have been instances of some being called out as AI generated video or work that took days and millions of tokens and not a one shot as claimed.

The first few days Astra was out, I attempted to reproduce a few of the demos I saw on twitter, and it was clear Astra had a distinctive style and certain limitations when working in short sessions with tools like Blender that could distinguish genuine demos from bs for retweets.

People putting out these demos should share their sessions to really show what was going on, and I invite people to test models and try to replicate what they see and draw their own conclusions.