Live data from Hacker News

Yt-dlp: Upcoming new requirements for YouTube downloads

github.com

471–480 of 635 posts

Re: Yt-dlp: Upcoming new requirements for YouTube downloads

#471

Earlier quoted context omitted.

I think you've missed the point entirely. The point is that if everyone is using a single browser (not just Chrome/Chromium) then that actor gets disproportionate control over the internet. That's not good for anyone. The specific gripe to Chromium is that _Google_ gets that say, and I think they are less trustworthy than other actors. I'm not asking anyone to trust Mozilla, but anyone suggesting Mozilla is less trus…

> I'm not sure what you're trying to tell me here. I'm telling you that Firefox is going to be out of business soon because users favor ad blocking and blocking trackers. That is the trend. Firefox isn't growing anymore. > Honestly I have more questions looking at the Brave "transparency" report, as it seems to have more information about users than Firefox... Metrics can be transmitted without revealing the user. Th…

  > I'm telling you that Firefox is going to be out of business soon
Do you not think everyone saying Firefox is going to be out of business soon plays a role in this?

Regardless, I think you've ignored the root of my argument. I'm not trying to be a Firefox fanboy here but it's not like there's many options. The playing field is Chrome, Firefox, Safari. So only one of these is not "big tech".

  > Metrics can be transmitted without revealing the user. This is well known.
This is not well known and I think you've kinda "told on yourself" here. It is fairly well known in the privacy community that it is difficult to transmit user data without accidentally revealing other information. Here's a rather famous example[0,1]. I'd encourage you to read it and think carefully about how deanonymization might be possible after just reading a description of the datasets they deanonymize.

  > You can't suggest anything. I am done with this conversation.
If you wish to disengage then that is your choice. I am really trying to engage with you faithfully here. I'm not even really attacking Brave here, as my critique is over the Chromium ecosystem. I think if you look at my points again you can see how they would dramatically shift if Brave were based off of Gecko or Webkit. Honestly, I would be encouraging Brave usage were it under those umbrella. Or even better, if it had its own engine! Because my point is about monopolization.

[0] https://courses.csail.mit.edu/6.857/2018/project/Archie-Gers...

[1] https://arxiv.org/abs/cs/0610105

Re: Yt-dlp: Upcoming new requirements for YouTube downloads

#472
post #391

Days of just getting data off the web are coming to an end as everything requires a full browser running thousands of lines of obfuscated js code now. So instead of a website giving me that 1kb json that could be cached now I start a full browser stack and transmit 10 megabytes through 100 requests, messing up your analytics and security profile and everyone's a loser. Yay.

It's an arms race. Websites have become stupidly/unnecessarily/hostilely complicated, but AI/LLMs have made it possible (though more expensive) to get whatever useful information exists out of them. Soon, LLMs will be able to complete any Captcha a human can within reasonable time. When that happens, the "analog hole" may be open permanently. If you can point a camera and a microphone at it, the AI will be able to ma…

The future will just be every web session gets tied to a real ID and if the service detects you as a bot you just get blocked by ID.

Re: Yt-dlp: Upcoming new requirements for YouTube downloads

#473

Earlier quoted context omitted.

I mean these numbers are just made up anyways, so why are engineers concerned with them? The idea of engineers needing to justify monetary value is just... ill conceived. They should be concerned with engineering problems. Let the engineering manager worry about the imaginary money numbers. > User retention is a thing. Problem is no one needs to care about the product's quality if the product has the market cornered.…

> I mean these numbers are just made up anyways, so why are engineers concerned with them? That's what they're directly or indirectly being graded on. Even if they don't have to show how their work impacted the company's bottom line, their managers or their managers' managers have to, and poop just rolls downhill. > The idea of engineers needing to justify monetary value is just... ill conceived. They should be conce…

  > That's what they're directly or indirectly being graded on.
I think you'd agree that this should have never been the case. Engineering managers or project managers, sure. But engineers? That's just silly.

We need firewalls. One group's primary concern needs to be on the product. Another group's primary concern needs to be on keeping the business alive and profitable.

Too much of the former and you fail to prioritize the right work. Too much of the latter and you build vaporware. The downsides of biasing in one direction is certainly worse than the other...

  > my wife hates that I see this in everything, but - hidden inflation is a thing.
Lol, your wife might have a field day with mine...

I have a fundamental belief that there's far more complexity than we let on. That as we advance complexity only increases. What was once rounding errors end up becoming major roadblocks. It's the double edged nature of success: the more you improve the harder it is to improve. I truly will never understand how everyone (including niche experts) thinks things are so simple.

But my partner is doing her PhD in economics, so she also thinks about opportunity costs quite a lot but I think she (and a lot of her friends) were quite unaware of how a lot of stuff operates in tech[0].

Probably doesn't help that, as you know, I'm not great at brevity :/

[0] My favorite thing to at her department get togethers (alcohol is always involved) is to introduce them to open source software. Quite a number of them find it difficult to understand how much of the world is based on this type of work and how little money it makes. Not to mention the motivations behind it. The xz hack led to some interesting discussions...

Re: Yt-dlp: Upcoming new requirements for YouTube downloads

#474

Earlier quoted context omitted.

> I'm not sure what you're trying to tell me here. I'm telling you that Firefox is going to be out of business soon because users favor ad blocking and blocking trackers. That is the trend. Firefox isn't growing anymore. > Honestly I have more questions looking at the Brave "transparency" report, as it seems to have more information about users than Firefox... Metrics can be transmitted without revealing the user. Th…

> I'm telling you that Firefox is going to be out of business soon Do you not think everyone saying Firefox is going to be out of business soon plays a role in this? Regardless, I think you've ignored the root of my argument. I'm not trying to be a Firefox fanboy here but it's not like there's many options. The playing field is Chrome, Firefox, Safari. So only one of these is not "big tech". > Metrics can be transmit…

[deleted]

Re: Yt-dlp: Upcoming new requirements for YouTube downloads

#475

Earlier quoted context omitted.

[flagged]

Yeah because everyone who has a user experience feedback about a piece of software is magically a skilled programmer? The smug "PRs accepted" doesn't help anyone. Expressing hope for a feature at least shows potential implementers that the feature is wanted.

ideas/assholes/everyone has one, etc

Re: Yt-dlp: Upcoming new requirements for YouTube downloads

#476

Earlier quoted context omitted.

Firefox is in decline and Brave will soon overtake it. Brave blocks ads natively. There is a lot of advantage in that but we also may eventually have a new model that funds the internet. And I don't see Firefox or Safari disrupting advertising. https://data.firefox.com/dashboard/user-activity https://brave.com/transparency/

I think you've missed the point entirely. The point is that if everyone is using a single browser (not just Chrome/Chromium) then that actor gets disproportionate control over the internet. That's not good for anyone. The specific gripe to Chromium is that _Google_ gets that say, and I think they are less trustworthy than other actors. I'm not asking anyone to trust Mozilla, but anyone suggesting Mozilla is less trus…

I'm a loyal Brave user and I feel my loyalties being swayed right now...

Re: Yt-dlp: Upcoming new requirements for YouTube downloads

#477

Earlier quoted context omitted.

Games Done Quick has raised 10s of millions for charity. I suspect they could raise a few thousand for a few dozen TB of nvme storage if they wanted to host a speedrun archive.

YouTube get 700,000 hours of video uploaded every day . That's 4.3 PB added per day. You may need more than a few dozen TB... https://www.reddit.com/r/AskProgramming/comments/vueyb9/how_...

That's because youtube allows almost everything sfw to be hosted on their platform and without any limits.

I can imagine if they've added rate-limiting, e.g. 30GB per IP per week - that would've reduced amount of crap, literal white noise and spam/scam videos uploaded to Youtube in several magnitudes. Another strategy is, if a video doesn't get 1000 views after a week - it's deleted.

Re: Yt-dlp: Upcoming new requirements for YouTube downloads

#478
post #94

Just the other day there was a story posted on hn[0][1] that said YouTube secretly wants downloaders to work. It's it's always been very apparent that YouTube are doing _just enough_ to stop downloads while also supporting a global audience of 3 billion users. If the world all had modern iPhones or Android devices you'd bet they'd straight up DRM all content [0] https://windowsread.me/p/best-youtube-downloaders [1] h…

More specifically, yt-dlp uses legacy API features supported for older smart TVs which don't receive software updates. Eventually once that traffic drops to near zero those features will go away.

So more people using yt-dlp will increase the likelihood of Youtube keeping legacy APIs?

Re: Yt-dlp: Upcoming new requirements for YouTube downloads

#479

Earlier quoted context omitted.

I was talking to "my friend" about how I'm annoyed my calendar duplicates holidays because it imports from multiple calendars and he asked me "what value" would be provided if this was solved. Confused I said it pushes things off so I can't read events. He clarified he meant monetary value... We're both programmers so we're both know we're talking about a one line regex... I know quite a number of people like this an…

> We're both programmers so we're both know we're talking about a one line regex... As a big tech programmer, it's almost never that simple... Small edges cases not covered by a one line regex can mean big issues at scale, especially when we're talking about removing things from a calendar.

  > As a big tech programmer, it's almost never that simple...
I'll be fair and agree that I'm being a bit facetious here. But let's also admit that if you are unable to dedupe entries in a calendar with identical names then something is fundamentally broken.

I did purposefully limit to holiday calendars as an example because this very narrow scope vastly simplifies the problem, yet is a real world example you yourself can verify.

You're right that edge cases can add immense complexities but can you really think of a reason it should be difficult to dedupe an event with identical naming and identical time entries, especially with the strong hint that these are holidays? Let's even just limit ourselves to holidays that exclusively fall over full day periods (such as Labor Day).

Do you really think we cannot write a quick solution that will cover these cases? The cases that dominate the problem? A solution whose failure mode results in the existing issue (having dupes)? Am I really missing edge cases which require significantly more complex solutions that would interfere with the handling of these exceptionally common cases? Because honestly, this appears like a standard table union problem. With the current result my choices are having triplicate entries, which has major consequences to usability, or the disabling of several calendars, which fails to generalize the problem and also results in missing some minor holidays. Honestly, the problem is so bad I'd be grateful even if I had to manually approve all such dedupes...

If not, I'd really like to hear. Because it really means I've greatly mischaracterized the problem and I should not be using this example. Nor the example of a failure to FIND contacts with identical names, nicknames, phone numbers, birthdays, and differ only on an email address and note entry. Because I have really been under the strong impression that the latter is a simple database query where we should return any entry containing matches (failure mode being presenting the user with too many matches rather than a lack of matches. We can sort by number of duplicate fields and display matches in batches if necessary. A cumbersome solution is better than the current state of things...).

I'm serious in my request but if I have made a gross mischaracterization then I think you'd understand how silly this all looks. I really do want to know because this is just baffling to me.

If I truly am being an idiot, please, I encourage you to treat me like one. But don't make me take it on your word.

Re: Yt-dlp: Upcoming new requirements for YouTube downloads

#480
post #391

Earlier quoted context omitted.

It's an arms race. Websites have become stupidly/unnecessarily/hostilely complicated, but AI/LLMs have made it possible (though more expensive) to get whatever useful information exists out of them. Soon, LLMs will be able to complete any Captcha a human can within reasonable time. When that happens, the "analog hole" may be open permanently. If you can point a camera and a microphone at it, the AI will be able to ma…

The future will just be every web session gets tied to a real ID and if the service detects you as a bot you just get blocked by ID.

I definitely agree logins will be required for many more sites, but how would the site be able to distinguish humans from bots controlling the browser? Captcha is almost obsolete. ARC AGI is too cumbersome for verifying every time.
Post reply on HN