These AI explainers are taking over. I tried getting this working back in May. The models were okay but not quite there and it was a ton of effort. Opus 5.5 seems like the tipping point.
That video in the README is amazing! Will definitely dig into this project.
We're not using Opus 5.5 btw, as it would be too expensive. Our goals is to eventually get the cost down to 1 cent for a 1 minute video, and with 1 second delay from submit to playback. For that, we need dirt cheap and lightning fast models.
honestly it's horrible compared to some of the Opus 5.5 ones. But yeah, at the latency/speed you're looking for it's a different ballpark. Let Opus write the style and reference screens/animations, let your small fast model assemble it from parts.
It did do what I expected from such a system:
- It removed the fun parts of the article and the way it is written.
- Stated some interpretations as fact without naming the source
Article '"That gunshot prematurely aged my brain,” she said'
Summary 'The impact prematurely aged her brain'
One surprise is did not like: The sharpest concrete thing was a .45 caliber bullet. Feels like some sort of attempt to add original humor and I do not like it.
This is the worst kind of article for this kind of summary and I feel like it really shows the most and least impressive parts of the system.
Thanks! Several people on our team are actually exactly like you: they don't really use videos themselves for learning, but see the utility for others, and enjoy the technical challenges around building a high-performant video format.
I can’t stand this Silicon Valley trend of having launch videos for features of 2 people sitting in front of a camera but it also shows the camera in the beginning to trick you into thinking “oh they do this so casually haha so fun, look how relatable they are”
I hate all of this - OpenAI, Claude, Cursor, Windsurf etc
It screams detached Silicon Valley tech company to me, and I will never watch a single of them
Indeed. In much the same vein, I simultaneously accept that:
1. I loathe video with the white-hot glow of a thousand angry suns, and loathe LLM slop-digests of things more still.
2. I spend tremendous amounts of time in situations where I can listen but not read, whether driving or on my bike, and this is a legitimately useful way to catch up on HN in those situations.
I hate this too, but I acknowledge some may find it useful.
This should not be judged as a way to read hn, where users reading, it's more of a demo for technical users of their core startup tech.
I still don't like it, I feel it's a very small script with like ffmpeg and an LLM HTTP call, but that same thing was said about dropbox. It's also like 2 years late, so the timing is off.
I'm often negative in the comments of people launching their startup ideas, but I think that's fair, I don't like the atmospheres where we are supposed to pep talk each other and tell ourselves that it's really cool.. and then their project dies while everyone pats them in the back, I think yc is explicitly about that type of nuanced feedback, this needs a pivot as-is.
This is actually quite impressive from an engineer viewpoint. I just have the feeling that the videos quickly become very monotonous and rather boring due to the monotonous AI voices. If somehow you could bring dynamic variation in these videos that would be fantastic.
Thanks! And I agree. We're getting tons of requests for more nuance in voice selection from our users. You're currently able to say i.e. "Australian English female" in your prompts today, but you should also ideally be able to describe the voice characteristic (i.e. like a funny grandpa, engaged news reporter).
What Gemini 3.8 Flash TTS is doing with generative voice design in this area super interesting.
Presumably you also like that they are explaining things right? I mean that seems like the more critical step. Otherwise if it's just Marquise Brownlee, you could just watch Marquise Brownlee say the same sequence of random words for 3 minutes a few times per day.
I think this is exactly the opposite of what AI should be used for. It is going to make people dumb.
Watching a video instead of engaging your own brain makes you feel like you learned something without actually trying and I would bet it works much less well.
If someone is there adding something beyond what is already there - ie someone like Maruqise. Then it makes sense for them to be there.
If not it is just brain rot.
People already have issues with concentration. Allowing them to further allow that muscle to atrophy will not be good.
All of what you said is great, and I especially detest the proliferation of slop videos on YouTube, where it makes no sense to drown out the ample supply of such personalities and insights with AI voices summarizing wikipedia or Reddit threads over AI imagery slideshows.
But given how good a job this seemed to do at giving me more than just headlines, I'd love to have maybe even just an audio podcast feed of the top 10 stories like this compiled a few times a day. I would listen while I'm doing things when reading isn't practical.
tl;dr agree that we don't need this to replace reading, but I see that it can be a useful tool.
I just had opus make me a tts mobile app that grabs the top 20 HN stories and the top 5 comments for each and read them out to me (using one of the higher quality built in TTS voices on my pixel). The summarization is the worst part of this imo, so pure TTS wins
Yeah I do see your point. And for some reason making it into audio does seem less offensive to me. Maybe it’s because it at least transforms it into something I can do while doing something else. Whereas a video is basically the same mode of operation: staring at a screen.
I’m still basically against this though. I consider it slop.
Yeah, there’s been an explosion of that guff over the past few weeks. “POV: You are a $x who $y”. “Every level of being a $x”. “$x explainer”.
The quality is getting very smooth but that just makes it worse because you can’t quite trust anything any more. It’s like the mental equivalent of HFCS.
This looks a good idea to let people choose their preferred medium. I am not into video either, but I turned one of my latest blog article into a video using Claude, since I thought it would be “easy.” I asked it to build the video by screencasting a browser presenting the interactive diagrams of the article and add some titles. It build some kind of recording app for me. It still took me many hours, but I think I would be faster if I redid it. I would pay to have it done faster, but I want to keep control of the text, not have a summary.
You can use MCP to generate an Explainer video next time if you want. It's currently free, as the tokens for directing and scripting it is offloaded to your agent. And then we cover the TTS for the time being.
I like video as a format because it's pattern-breaking. Consuming thousands of very similarly produced videos sounds even more depressing than consuming LLM blogs and articles. EDIT: That aside, this is of course extremely impressive.
Great. Wordpress plugin would work. Get publishers to embed these videos at the top of their own articles. Allow them to tweak/fix/fine-tune them. Accept micropayments and on-demand generation. Readers can pay for better videos or "deep-dives."
This will work for young readers too, illiterates, sensory-impaired and foreigners. All four!!
This is actually really useful to me, thanks! I often skip reading interesting articles because it's either too technical or because my eyes have had enough for the day. Listening to a summary video like this is a great solution!
Edit: I think I'd prefer this website to be audio only. I don't appreciate the AI generated imagery, because of how poorly it represent the content it's describing.
I like the concept and these explainers. It sounds like this is using a HTML based format to be similar to like a Flash / Shockwave animation to keep data size down.
I wonder what the challenge would be of making these videos more interactive would be? Also I do wonder if there is any research happening in making AI explainers more trustable.
Btw, we are hiring a new developer to join us in Oslo, Norway. We pay well and give a lot of freedom wrt working from home or at the office (hybrid solution). If getting HTML-based video generation down to 1 cent for 1 minute with 1 second delay, then please email me: per@scrimba.com.
We currently spend ~$0.4 on a video (without images) and a have a few seconds of delay before playback. So it's a matter of turning every stone in order to optimize it further. We're going to add Jev asap, as it's an obvious improvement along both axes.
Nice concept, but why not generate one video per post rather than re-generating the video for each click per user? The result is going to be the same, and it will be significantly cheaper to run.
Don't listen to the naysayers, please. This is both impressive _and_ useful.
As someone with limited time (the only one with such an issue in this comment section, it seems), I really appreciate having a quick video to get the gist of a link in under two minutes. It helps me decide whether it's worth diving deeper.
Currently, I rely on two main signals when choosing what to read: 1) direct interest in the topic, and 2) upvote/comment counts if the topic is new to me.
The downside is obvious: I risk missing out on hidden gems that I’d love, but that lack those initial signals.
Having summary videos lowers the cost of exploration (at least on my side) and gives me an easy way to broaden my horizon.
This feels like when a decent book is made into a movie, and then the movie does well so someone will take the movie and summarize the plot in a YouTube video. TLDR: this is TLDR for TLDR. Just read more from the source.
Yeah I feel like this would be better as summaries of textbook chapters or information rich sources, a lot of HN posts are already basically text summaries of a complex topic that it doesn't really make sense to further summarize them.
To me, the use case is to go from an often opaque headline ("Using Nix and containerd with Jev on macOS") to a paragraph which hopefully gives some strong hints as to what those things are and why it's interesting, because the articles often are written for an audience already deeply into all the topics. Because often a few of those terms I might have no clue about and thus scroll on by, but with a brief "why you should care" explanation, I might realize it's actually something cool.
This is great - many of us who are technical but not working in hardcore tech probably scroll by quickly on stuff we have never heard of. Getting a quick synopsis like this may drive more traffic and understanding.
Clicked on the demo "What caused the fall of the Roman Empire?" and immediately saw the same inadequacies I've seen in every other system like this: any attempt at visualizing a point ends up being absolute nonsense. Graphics and charts that confuse more than help, or even contradict what is being said.
It ends up being a hurdle I don't think these services can surpass, given these systems understand nothing, and have no spatial reasoning. Any of these systems that claim to help explain things visually end up being a huge liability for the learner and I would never recommend them to someone.
They understand a heck of a lot more than they did. We’ve gone from “understands how to form a coherent sentence, mostly” to “understands enough about the world to generate a 20 minute infotainment video that you wouldn’t pick as AI unless you’re looking for the current tells” in, what, 5 years?
It’s worth pondering what the world would be like if current AI methods don’t hit a glass ceiling, and get to the point where they can legitimately produce better, more creative, higher ‘nutritional value’ output than humans can. How do we spend our lives?
That was one of the challenges when building this. We ended up using Firecrawl. But even so there are a few pages we aren't able to read. In which case we use the Algolia API to get the HN comments and generate the video about the HN reaction on the article instead.
Instead of using HN comments which can be a gross distortion of an article, it is better to ask the user to upload a PDF of the full text of the article. It is easy for the user to create such a PDF. Thereafter, subsequent users should have an opportunity to upload a new PDF if they so desire.
I just tried it out - it's pretty cool. The postscript was actually noted in the generated video about this topic, which I find amusing.
I have to say, if HN were to add an AI generated paragraph summary about each link at the top of each comments page, it would probably go a long way to improving the discussions about a lot of topics. The number of people who comment based on the title of a link alone is pretty high, and I'd bet the vast majority of commenters only skim the linked pages anyways. Might as well get everyone on the same page with a quick overview.
For most articles these days, we'd have to first have a separate bot get the archive.is link for it!
But I'm skeptical that people would embrace it. It's more about the optics, and it also relates to the old principle of "Read the article -- don't start opining based only on the headline." Many would probably say that you shouldn't offer an opinion if you've only read an AI distillation.
It could be wrong - or more likely, the article acknowledges likely objections and refutes them well, but that part didn't make it into the summary, so everyone starts raising very un-insightful points as though the author was oblivious to them.
I don't dislike that it's video (because I know people who won't read can't be "made to read" by there not being "nice enough explainer videos"), but this is the kind of stuff that bothers me in both video and text, and greatly. I don't want to rant about it here because it's not specific to your product at all, it's not even specific to LLM, the internet and especially youtube is full of stuff that neither speakers/authors nor audience seem to ever parse. And if you're just downstream of big models you probably can't really influence that. But since you are NIH enjoyers, maybe you could train your own at some point? I don't want to consume such videos, but I do want those who do to have nicer ones, because that'd be good for me, too.
We are not particularly pleased with the state of our visualizations, so I totally agree. If you have any examples of how you'd like to see the visualizations, please share them!
It’s still a big thing to shop, iteration is the gift.
I think visualizations can be inspirational but often need to be personalized. Whether it’s audience, content, or presentation style the more I watched the HN feed, the more I wanted to be able to adjust and tweak certain things.
It was less about video at one point I could even drive and listen as an a podcast.
We've always gotten a ton of positive comments from people with ADHD who've used our HTML videos to learn to code. Despite it not being in our target group at all when we started out.
I gave it a fair shake (10 or so).
I'll say this... it does the job of "I can't spend 45 minutes on a journal article that's already over my head, or get through 20 pages of blog slop, so please give me the summary version of this".
The select comment/conversation slide is a nice touch, though it's sure making some choices there.
Main problem seems to be that the sameness of the format (it's clearly following a strict template) and voice do make it feel very repetitive, real fast. I'm not sure how you get consistent results AND solve for that, though.
I would stick with this. There's huge potential here. I'm not going to lie, I did a video of this, and it was pretty slop. Some of it didn't even make sense and was super irrelevant. That being said, refined, there's really actually some major potential with this idea.
These AI explainers are taking over. I tried getting this working back in May. The models were okay but not quite there and it was a ton of effort. Opus 5.5 seems like the tipping point.
I made an OSS framework for these for when you want to go beyond one-shoting it: https://github.com/scosman/videowright
- Voiceovers: aligns animations to the voiceover, can generate voiceover with elevenlabs, or will transcribe and timestamp a real voiceover
- can reorder scenes both in code, and using ffmpeg for audio.
- interactive controls during authoring, can ask for micro edits or re-builds
- MP4 export/encoder
- Generates a video from a prompt (obvs)
That video in the README is amazing! Will definitely dig into this project.
We're not using Opus 5.5 btw, as it would be too expensive. Our goals is to eventually get the cost down to 1 cent for a 1 minute video, and with 1 second delay from submit to playback. For that, we need dirt cheap and lightning fast models.
honestly it's horrible compared to some of the Opus 5.5 ones. But yeah, at the latency/speed you're looking for it's a different ballpark. Let Opus write the style and reference screens/animations, let your small fast model assemble it from parts.
Could you link great opus 5.5 ones?
https://news.ycombinator.com/item?id=49836374
Main post and comments have a few
I really dislike this but have to say it is impressive.
To test it I tried "Who Killed Paulina Borsook's Career?" The summary is correct I did not hear any errors. https://www.wired.com/story/paulina-borsook-profile/ https://hn.watch/?item=49866903
It did do what I expected from such a system: - It removed the fun parts of the article and the way it is written. - Stated some interpretations as fact without naming the source Article '"That gunshot prematurely aged my brain,” she said' Summary 'The impact prematurely aged her brain'
One surprise is did not like: The sharpest concrete thing was a .45 caliber bullet. Feels like some sort of attempt to add original humor and I do not like it.
This is the worst kind of article for this kind of summary and I feel like it really shows the most and least impressive parts of the system.
Because we all contain multitudes, I can simultaneously accept that:
1. I hate everything about this because I vastly prefer text over video for the same content, especially AI generated video
2. There are a lot of people for whom video is their preferred medium and so this will be valuable to them.
It's a technically cool project and your cost-per-video is impressively low. Best of luck!
Thanks! Several people on our team are actually exactly like you: they don't really use videos themselves for learning, but see the utility for others, and enjoy the technical challenges around building a high-performant video format.
I can’t stand this Silicon Valley trend of having launch videos for features of 2 people sitting in front of a camera but it also shows the camera in the beginning to trick you into thinking “oh they do this so casually haha so fun, look how relatable they are”
I hate all of this - OpenAI, Claude, Cursor, Windsurf etc
It screams detached Silicon Valley tech company to me, and I will never watch a single of them
Indeed. In much the same vein, I simultaneously accept that:
1. I loathe video with the white-hot glow of a thousand angry suns, and loathe LLM slop-digests of things more still.
2. I spend tremendous amounts of time in situations where I can listen but not read, whether driving or on my bike, and this is a legitimately useful way to catch up on HN in those situations.
Is it a legit summary? Does it derail you?
I concur with only #1.
I have ocd so I concur with only #2 for equilibrium.
My my first comment is still -1. Everything is out of balance still.
I hate this too, but I acknowledge some may find it useful.
This should not be judged as a way to read hn, where users reading, it's more of a demo for technical users of their core startup tech.
I still don't like it, I feel it's a very small script with like ffmpeg and an LLM HTTP call, but that same thing was said about dropbox. It's also like 2 years late, so the timing is off.
I'm often negative in the comments of people launching their startup ideas, but I think that's fair, I don't like the atmospheres where we are supposed to pep talk each other and tell ourselves that it's really cool.. and then their project dies while everyone pats them in the back, I think yc is explicitly about that type of nuanced feedback, this needs a pivot as-is.
This is actually quite impressive from an engineer viewpoint. I just have the feeling that the videos quickly become very monotonous and rather boring due to the monotonous AI voices. If somehow you could bring dynamic variation in these videos that would be fantastic.
Thanks! And I agree. We're getting tons of requests for more nuance in voice selection from our users. You're currently able to say i.e. "Australian English female" in your prompts today, but you should also ideally be able to describe the voice characteristic (i.e. like a funny grandpa, engaged news reporter).
What Gemini 3.8 Flash TTS is doing with generative voice design in this area super interesting.
The thing I like about an explainer video is the personality and insight of the person giving it.
Like Marquise Brownlee just has opinions I care about and I watch his videos for this reason.
If a video just explains something I could just read all I’m getting is a layer of obfuscation.
Presumably you also like that they are explaining things right? I mean that seems like the more critical step. Otherwise if it's just Marquise Brownlee, you could just watch Marquise Brownlee say the same sequence of random words for 3 minutes a few times per day.
I feel I need to explain beyond just a whine.
I think this is exactly the opposite of what AI should be used for. It is going to make people dumb.
Watching a video instead of engaging your own brain makes you feel like you learned something without actually trying and I would bet it works much less well.
If someone is there adding something beyond what is already there - ie someone like Maruqise. Then it makes sense for them to be there.
If not it is just brain rot.
People already have issues with concentration. Allowing them to further allow that muscle to atrophy will not be good.
All of what you said is great, and I especially detest the proliferation of slop videos on YouTube, where it makes no sense to drown out the ample supply of such personalities and insights with AI voices summarizing wikipedia or Reddit threads over AI imagery slideshows.
But given how good a job this seemed to do at giving me more than just headlines, I'd love to have maybe even just an audio podcast feed of the top 10 stories like this compiled a few times a day. I would listen while I'm doing things when reading isn't practical.
tl;dr agree that we don't need this to replace reading, but I see that it can be a useful tool.
I just had opus make me a tts mobile app that grabs the top 20 HN stories and the top 5 comments for each and read them out to me (using one of the higher quality built in TTS voices on my pixel). The summarization is the worst part of this imo, so pure TTS wins
Yeah I do see your point. And for some reason making it into audio does seem less offensive to me. Maybe it’s because it at least transforms it into something I can do while doing something else. Whereas a video is basically the same mode of operation: staring at a screen.
I’m still basically against this though. I consider it slop.
Yeah, there’s been an explosion of that guff over the past few weeks. “POV: You are a $x who $y”. “Every level of being a $x”. “$x explainer”.
The quality is getting very smooth but that just makes it worse because you can’t quite trust anything any more. It’s like the mental equivalent of HFCS.
0:04 for the first , 0:02 for the second. I'm personally all done with that.
Please, nobody click on the link to this hn item within hn.watch as it would be even more dangerous than typing 'google' into Google
I found that video to be one of the clearest. TBH they are all pretty good.
This is pretty crazy. It's not hard to imagine something like Reddit deploying this as a first party feature.
This looks a good idea to let people choose their preferred medium. I am not into video either, but I turned one of my latest blog article into a video using Claude, since I thought it would be “easy.” I asked it to build the video by screencasting a browser presenting the interactive diagrams of the article and add some titles. It build some kind of recording app for me. It still took me many hours, but I think I would be faster if I redid it. I would pay to have it done faster, but I want to keep control of the text, not have a summary.
The article: https://vincent.bernat.ch/en/blog/2026-spanning-tree-video
The tool built by Claude to make the video: https://github.com/vincentbernat/vincent.bernat.ch/tree/2808...
Nice! Will check out that repo.
You can use MCP to generate an Explainer video next time if you want. It's currently free, as the tokens for directing and scripting it is offloaded to your agent. And then we cover the TTS for the time being.
I like video as a format because it's pattern-breaking. Consuming thousands of very similarly produced videos sounds even more depressing than consuming LLM blogs and articles. EDIT: That aside, this is of course extremely impressive.
I don't really see why a new programming language is needed to make explainer pages (they are not actually videos).
Seems like over engineering - why not just make JavaScript/HTML?
This is super cool.
Not affiliated but trymyrepo.com is very neat as well. Seems similar space.
Also I had Muse agent create a skill that can do videos. Worked pretty well. Things are getting wild.
Lastly, this might fun to those who like this space: (my thing)
https://github.com/jgbrwn/mst3k-anything
Are you planning to consume article comments for your videos?
Great. Wordpress plugin would work. Get publishers to embed these videos at the top of their own articles. Allow them to tweak/fix/fine-tune them. Accept micropayments and on-demand generation. Readers can pay for better videos or "deep-dives."
This will work for young readers too, illiterates, sensory-impaired and foreigners. All four!!
This is actually really useful to me, thanks! I often skip reading interesting articles because it's either too technical or because my eyes have had enough for the day. Listening to a summary video like this is a great solution!
Edit: I think I'd prefer this website to be audio only. I don't appreciate the AI generated imagery, because of how poorly it represent the content it's describing.
I like the concept and these explainers. It sounds like this is using a HTML based format to be similar to like a Flash / Shockwave animation to keep data size down.
I wonder what the challenge would be of making these videos more interactive would be? Also I do wonder if there is any research happening in making AI explainers more trustable.
you know what might be interesting? The "HN 10 minute" -- 20 seconds per front page hn story as a podcast, daily.
Yes! We can set that up as a video quite easily. Don't have the infrastructure to push it as a podcast yet though.
just ffmpeg it -vn and post it as an rss xml ... the audio sans video is a fine first step.
Post them to YouTube as shorts and capture view revenue.
Btw, we are hiring a new developer to join us in Oslo, Norway. We pay well and give a lot of freedom wrt working from home or at the office (hybrid solution). If getting HTML-based video generation down to 1 cent for 1 minute with 1 second delay, then please email me: per@scrimba.com.
>> getting HTML-based video generation down to 1 cent for 1 minute with 1 second delay
How would you do that?
We currently spend ~$0.4 on a video (without images) and a have a few seconds of delay before playback. So it's a matter of turning every stone in order to optimize it further. We're going to add Jev asap, as it's an obvious improvement along both axes.
You guys are going to hate this but next up: Each thread in a 10 second explainer and it just autoplays
And there’s no way to stop it
And one has no eyelids. And I must blink. And not topple over. And out.
Nice concept, but why not generate one video per post rather than re-generating the video for each click per user? The result is going to be the same, and it will be significantly cheaper to run.
We do cache them like that! As I wrote:
"We create them on-the-fly the first time someone clicks on a link."
Though I realise it could have been said more explicitly. But the second time someone clicks the link we re-use the same explainer.
Haha, I should have read more thoroughly! Thanks
Don't listen to the naysayers, please. This is both impressive _and_ useful.
As someone with limited time (the only one with such an issue in this comment section, it seems), I really appreciate having a quick video to get the gist of a link in under two minutes. It helps me decide whether it's worth diving deeper.
Currently, I rely on two main signals when choosing what to read: 1) direct interest in the topic, and 2) upvote/comment counts if the topic is new to me.
The downside is obvious: I risk missing out on hidden gems that I’d love, but that lack those initial signals.
Having summary videos lowers the cost of exploration (at least on my side) and gives me an easy way to broaden my horizon.
As a non-technical and oft confused reader, this is very helpful! Also a very happy Scrimba user :)
loved scrimba. thanks for building that. was anout to pushback bc of the torrent of show hn sloppy pists but this one is fun.
Thank you!
You should talk to the Editory guys, they did this for local news.
I'm an instructional designer, and I think this is awesome!
I'm going to be looking to reproduce this.
We have a free MCP so you are free too use our tool if you'd like to use it in your courses as well :)
Thanks!
It's actually very good...
I can see vertical versions of these doing well on YT shorts or TikTok.
I hope so! We actually already support vertical videos when generated via the web ui (by clicking the mobile button).
Did not think I would like that, but really good, simple and easy summary.
Why not HN.tv?
I like it! I bookmarked it. But I wouldn't pay for it.
This feels more useful as a gateway to reading than a replacement for it.
This feels like when a decent book is made into a movie, and then the movie does well so someone will take the movie and summarize the plot in a YouTube video. TLDR: this is TLDR for TLDR. Just read more from the source.
Yeah I feel like this would be better as summaries of textbook chapters or information rich sources, a lot of HN posts are already basically text summaries of a complex topic that it doesn't really make sense to further summarize them.
To me, the use case is to go from an often opaque headline ("Using Nix and containerd with Jev on macOS") to a paragraph which hopefully gives some strong hints as to what those things are and why it's interesting, because the articles often are written for an audience already deeply into all the topics. Because often a few of those terms I might have no clue about and thus scroll on by, but with a brief "why you should care" explanation, I might realize it's actually something cool.
How does this capture the video?
It doesn't actually capture it as a video, it's just HTML with a voice over essentially.
Love this! I've already used it a bunch to go through some HN posts that I might not have otherwise read.
It's definitely a form of TLDR, but for watching?!
This feels like some kind of parallel to Jev in that. If it is super fast, what other use cases may be unlocked it, as you suggest.
Also, I love the simple, easy to remember domain name, HN.watch!
Glad to hear that. Adding Jev as our "pre-planner" is very high on our todo list. It's the perfect tool for this kind of thing.
This is great - many of us who are technical but not working in hardcore tech probably scroll by quickly on stuff we have never heard of. Getting a quick synopsis like this may drive more traffic and understanding.
Thanks!
infinite token glitch
Clicked on the demo "What caused the fall of the Roman Empire?" and immediately saw the same inadequacies I've seen in every other system like this: any attempt at visualizing a point ends up being absolute nonsense. Graphics and charts that confuse more than help, or even contradict what is being said.
It ends up being a hurdle I don't think these services can surpass, given these systems understand nothing, and have no spatial reasoning. Any of these systems that claim to help explain things visually end up being a huge liability for the learner and I would never recommend them to someone.
They understand a heck of a lot more than they did. We’ve gone from “understands how to form a coherent sentence, mostly” to “understands enough about the world to generate a 20 minute infotainment video that you wouldn’t pick as AI unless you’re looking for the current tells” in, what, 5 years?
It’s worth pondering what the world would be like if current AI methods don’t hit a glass ceiling, and get to the point where they can legitimately produce better, more creative, higher ‘nutritional value’ output than humans can. How do we spend our lives?
this is really cool and yet i absolutely hate it. ggs
it seems like with every passing year we're progressing towards making Wall-E a reality and idk how to feel about that
I don't see an RSS feed.
How does this even reliably read the webpages, since so many of them block bots?
That was one of the challenges when building this. We ended up using Firecrawl. But even so there are a few pages we aren't able to read. In which case we use the Algolia API to get the HN comments and generate the video about the HN reaction on the article instead.
Instead of using HN comments which can be a gross distortion of an article, it is better to ask the user to upload a PDF of the full text of the article. It is easy for the user to create such a PDF. Thereafter, subsequent users should have an opportunity to upload a new PDF if they so desire.
I just tried it out - it's pretty cool. The postscript was actually noted in the generated video about this topic, which I find amusing.
I have to say, if HN were to add an AI generated paragraph summary about each link at the top of each comments page, it would probably go a long way to improving the discussions about a lot of topics. The number of people who comment based on the title of a link alone is pretty high, and I'd bet the vast majority of commenters only skim the linked pages anyways. Might as well get everyone on the same page with a quick overview.
For most articles these days, we'd have to first have a separate bot get the archive.is link for it!
But I'm skeptical that people would embrace it. It's more about the optics, and it also relates to the old principle of "Read the article -- don't start opining based only on the headline." Many would probably say that you shouldn't offer an opinion if you've only read an AI distillation.
It could be wrong - or more likely, the article acknowledges likely objections and refutes them well, but that part didn't make it into the summary, so everyone starts raising very un-insightful points as though the author was oblivious to them.
> And finally, a real pixel-based video of the tool: https://www.youtube.com/watch?v=k6rbHmBxSEs
"a compressor uses three main organs" @ 2:97
I don't dislike that it's video (because I know people who won't read can't be "made to read" by there not being "nice enough explainer videos"), but this is the kind of stuff that bothers me in both video and text, and greatly. I don't want to rant about it here because it's not specific to your product at all, it's not even specific to LLM, the internet and especially youtube is full of stuff that neither speakers/authors nor audience seem to ever parse. And if you're just downstream of big models you probably can't really influence that. But since you are NIH enjoyers, maybe you could train your own at some point? I don't want to consume such videos, but I do want those who do to have nicer ones, because that'd be good for me, too.
This is pretty cool, I'd want to see the visualizations a little different but that's a personal preference.
We are not particularly pleased with the state of our visualizations, so I totally agree. If you have any examples of how you'd like to see the visualizations, please share them!
It’s still a big thing to shop, iteration is the gift.
I think visualizations can be inspirational but often need to be personalized. Whether it’s audience, content, or presentation style the more I watched the HN feed, the more I wanted to be able to adjust and tweak certain things.
It was less about video at one point I could even drive and listen as an a podcast.
Thanks, I hate it.
Kidding, I already sent it to a friend with ADHD who has been struggling to remain anchored to the tech world in any way besides Shorts.
But, culturally, I definitely hate the trend it implies!
Neat. Thank you for sharing.
Hahah! Thanks.
We've always gotten a ton of positive comments from people with ADHD who've used our HTML videos to learn to code. Despite it not being in our target group at all when we started out.
Hope your friend finds it helpful!
I gave it a fair shake (10 or so). I'll say this... it does the job of "I can't spend 45 minutes on a journal article that's already over my head, or get through 20 pages of blog slop, so please give me the summary version of this".
The select comment/conversation slide is a nice touch, though it's sure making some choices there.
Main problem seems to be that the sameness of the format (it's clearly following a strict template) and voice do make it feel very repetitive, real fast. I'm not sure how you get consistent results AND solve for that, though.
Leaning into brainrot is not going to unrot your brain, even - especially - if you have ADHD.
I watched the videos and confirmed they are worse than brainrot.
Now I want to watch the comments as a video too.
We usually include a section about the comments in the end of the video. But perhaps we should add a dedicated "Comments Video" as well.
yo this is wild
I would stick with this. There's huge potential here. I'm not going to lie, I did a video of this, and it was pretty slop. Some of it didn't even make sense and was super irrelevant. That being said, refined, there's really actually some major potential with this idea.
Thanks. Definitely lots of work left to make this feel non-AI-slop.
Think there's a ton of potential here too. So we're staying on course.
This is awesome!