Hacker Newsnew | past | comments | ask | show | jobs | submit | CamperBob2's commentslogin

I sure hope humans are in the loop at those companies, reviewing pics that aren't recognized as faces. If I ever see one of those captchas, they'll get to see some things that no human has ever, or should ever, see.

Nah, I just have my clanker deal with it. It's better at those stupid puzzles than I am.

Cloudflare doesn't do puzzles. It either says you're a bot or you're not a bot.

Patchright is a patched playwright that clears turnstile and other similar automation hostilities

Love when something decides I'm a bot and the decision is final. (Looking at you, CPU-World -- it even lets me know my IP address has been banned for "abuse", too.)

Right. Because IP addresses uniquely identify people. /eyeroll

Even if I were behind CGNAT (which I'm not), you don't see that message if you're already blocked. They won't even serve you the signup page then.

I remember being blown away when a then-unreleased version of GPT 5 took gold at the International Math Olympiad. Now I can run a model at home that can do that. We are more fortunate to have these tools than almost anyone is willing to acknowledge.

If it's vibecoded, who GAF what language is used?

This argument applies regardless of how you feel about vibecoding. Lately people have been asking for x64 machine language and getting reasonable results.


When they started leaving notes for their future selves, that rationale became a little more interesting. We're seeing the first stirrings of object permanence.

How many user pictures are there to choose from, though? This is a nice, elegant way to sample from the set, but the nature, size, and frequency of the problem don't justify more than five or ten seconds' or thought and one or two minutes of coding.

But multiplied by the number of times it'll run on the planet, and you have a surprisingly big impact.

Exactly this, if you count the amount of time the windows 11 context menu needs to pop up and then the second click to get to the old context menu across the entire globe for a month, you would get an insane amount of time and cycles wasted

I guess I'd do it once, the slow and easy way, and cache the result. But that's just me.

The billions of times aren't repeated times for the same person, it's because it's been done for the first time billions of times. This is the "choose a user's initial profile picture". They're not randomly changing it.

If its weights are open, that covers a multitude of other sins. Sufficiently-strong performance on the part of the new model would justify adapting existing tools to work with it.

What's a concrete example of how FORTRAN compilers are inherently better at this type of task than modern C/C++ or Rust compilers?

It's not really about the compilers ever since C99's restrict (you can almost always equal fortran performance now), it's how much more you have to think and validate about how to write the C/C++ plus the compiler and linker options to use.

Also with C99 you get many of the convenient aspects of fortran (<tgmath.h> and vardim array function args) and with C11 you also get <complex.h> now.

But you are still missing fortran's ** operator and it's more than just syntactic sugar, it accepts more types.


There are similar exceptions. For instance, the copy of licensed content that your computer makes in RAM in order to play it is explicitly exempt from being treated as a copyright violation.

News flash: people who scraped the Internet without permission to build their product complain when something vaguely similar is done to them. Water still wet, sky still blue. Film at 11.

(slibhb: Don't get me wrong, I agree with you 99%. But the frontier labs have zero moral authority here.)


As I understand, they got paid for the traces unlike owners of scraped websites. They sell text generation tool, so what's the problem if someone generates texts using it?

I'd send them the worlds smallest violin but Rufus is getting in the way of me finding it.

I had to install an add-on in Waterfox to stop Rufus from following me around and interjecting every 2 minutes.

I'd already stopped using Amazon for geopolitical reasons but I needed to get something in an emergency the last week (family member in the hospital, so I bent the rule) first time I'd seen Rufus, even if I wasn't boycotting Amazon for other reasons that monstrosity would have made me consider it.

I would not be surprised in the slightest if we later find out they are running those same open models to find useful traces or bits to incorporate into their own training. Lots of rules for thee but not for me from Big Ai

I look forward to a day when open models are so dominant that we stop considering traces to be some form of intellectual property that must be hidden from / manipulated for paying users.

It's that manipulation of inputs and outputs that really rubs me the wrong way


They are copying useful parts of open models 100% especially from deepseek.

Evidence?

not just the internet, but every commercially published written work in existence, and I doubt their highly publicized destructive scanning thing had managed to legitimize even a fraction of a percent.

this what is permissible for Jupiter is not permissible for a cow bullshit alone should tell people all they need to know about what kind of greasy sociopaths run "open"ai and (mis)anthropic, and how seriously you should take their purported stances on "safety" and other self-serving shit.


I understand people just get off posting stuff like this. But creating LLMs from the entire corpus of human text was a huge achievement. Distilling those models is much less of an achievement. It means China is further behind than we thought.

They're using competitor output as additional training input. Shrug.. It's not like they have access to the weights.

First movers rarely take the prize, though, do they? Les Paul and Mary Ford pioneered overdubbing voices back in the 1940s and 50s. The Beatles stole it from Buddy Holly. And Elton John from the Beatles.

Distilling is a massive achievement. I can run Qwen. I can't run GPT (TM). It's not a matter of X is better than Y. It's a matter of Y exists, X does not.

How is China further behind if distillation cannot stop? I think it's a reasonable strategy to follow, even if they could train from scratch.

distillation results in a worse product than the actual teacher model iirc

Nobody care if Chinese models are only 99%, 95% or 90% as good as SotA US models.

Because we only have weights and able to self-host Chinese ones. Gemma 4 and GPT OSS are nice to have, but nowhere close to that.


For some reason, what China is doing seems worse. Part of it is that I want the US to stay ahead of China.

Guidelines | FAQ | Lists | API | Security | Legal | Apply to YC | Contact

Search: