Hacker Newsnew | past | comments | ask | show | jobs | submit | igleria's commentslogin

v4 pro was decent then a better cheaper faster model comes now?

As a consumer I feel like hansel and gretel combined, deepseek could be the witch.


It's not unprecedented given that GLM 5.3 Flash was better and cheaper than GLM 5.2.

v4 flash has been working quite well for the majority of my personal projects, with occasional v4 pro or Kimi 3 for the most complicated tasks or to check the overall project progress (when vibe coding).

I must be doing something wrong. I gave v4 pro a try a couple of days ago, gave it a simple prompt like "clean up functions x and y in file z" and it would always start off promising, just to quickly get sidetracked, start hallucinating problems in the code, and just get stuck for hours until I interrupt it:

— hmm — 0x2D696370 — little-endian bytes: 70 63 69 2D = 'p','c','i','-' — hmm — WAIT — WAIT — !!!!! — *WAIT — WAIT — WAIT — WAIT — WAIT — WAIT — *HOLD ON — HOLD ON — HOLD ON — WAIT — WAIT — WAIT — WAIT — WAIT — WAIT — WAIT — !!!!!!!! — *WAIT — WAIT — WAIT — WAIT — WAIT — WAIT — WAIT — WAIT — *OK — WAIT — I THINK I FINALLY SEE THE WHOLE PICTURE — I NEVER READ IT — AND — THE LAYOUT — hmm — !!!!! — *WAIT — WAIT — WAIT — WAIT — WAIT — WAIT — WAIT — WAIT — WAIT — HOLD ON — HOLD ON — HOLD ON — HOLD ON

Then gave the same to Sonnet 5 and it was done 15 - 30 minutes later. I tried v4 pro both in claude code and codewhale with similar results. Haven't tried the new deepseek harness.


I've been using the 0731 Flash V4 model, via Opencode, and I've had no major issues. It feels very comparable to Opus 4.6/4.7, that I use at work. I haven't ran into any of the problem you mention, so that might be a quirk of V4 pro, the specific harness, or maybe the host you're using?

I used flash with pi and it worked pretty well.

It built this whole IaC plugin from scratch: https://github.com/fllstck/nebius-alchemy


I guess this was a heavily quantized version from openrouter? I've never had that experience in the last months of quite intensive use of the official deepseek api.

No, that was directly from deepseek. And it happened very consistently (3 or 4 times in a row, clean session every time, and across 2 harnesses). Guess I'll give it another try when v4.1 flash is available.

> Your local trailer park was never going to be able to afford sponsoring high energy

taxes.


It's funny that when faced with a weapon that could finish the war, option b was still selected (thanks Teller).

But the people faced with decisions close to IPO, well, we know what is happening.


It's a bit depressing. I personally try to enjoy the present with my loved ones as much as I can, the future is a bit impossible to forecast right now. I'm focusing on staying alive, mortgage payments, etc.

I share your concern...


If I was a company with a zero data retention contract involving OAI I would be asking for a third party audit of such claim of zero retention like, yesterday.

Could they say they don't retain, but do something "transformative" like use their own AI to summarize and paraphrase user sessions?

data collection companies regularly fuzz and mask data and call it a day. the fuzz and the mask quality is debatable.

Yes, and that’s exactly what I believe they are doing

Is there an implication of violation of ZDR here? Not a challenge. Just a request for clarification.

to my knowledge the mathematicians did not have ZDR so it would be incorrect to assume OAI violated such a thing.

I'm suggesting audits, not suing... if that is the implication.


By the way, the company that made it's entire product off of stealing all data it could get it's hand on while violating copyright and pirating, is not all of a sudden going to respect your data. If you think OpenAI or any major AI lab is going to give you true ZDR, I have a bridge to sell you.

So use bedrock or vertex or whatever. Those are the ZDR offerings. Or was it your intention to insinuate that the major cloud providers are conspiring with openai to violate their contractual obligations to their customers?

Yes. You're naive if you think any of these cloud providers care about your data when they're all in the midst of a AI revolution psychosis. They dont care about their reputation or what you think of them, they think they're going to have a machine god their side.

If my company finds any evidence of OpenAI violating ZDR, we'll sue for breach of contract and fraud, and collect damages. I think we'll be able to afford the bridge you're selling. You've got the title and title insurance, right?

Sure you will little bro. You signed away your right to arbitration a long time ago.

You realize OpenAI hired Apple employees and covertly had them stay working at Apple to steal from them. You think they're scared of your lawyers lol?


I'm certain my company didn't agree to arbitration, big bro. Our lawyers are putting the fries in the bag.

Isn't OpenAI being sued by Apple for their little stunt?

If you're saying "the big bad guys always win anon, just take the black pill," then there are tons of counterexamples. Remember Uber paying Google a sweet Bil for pulling this same trick with LIDAR firmware?


It can still be academic plagiarism even if they ticked the box to allow training on their prompts.

The idea OpenAI or Anthropic won't train on your data—even with an enterprise contract—is a fantasy at best, and dilusion at worst.

This is how every conspiracy theorist thinks: my enemy is Bad, and if they did a Bad thing, it would be Good for them, therefore they obviously did it. No evidence needed other than "motive" + my enemy is evil. But even if your enemy is evil, in this case, they would be fools to take the legal risk of violating their contract for the minimal upside of a tiny bit more training data (and fools to assume this would not be exposed in a large organization). So you need to assume your enemy is both evil and remarkably stupid.

I think it’s probably not surprising that they would go up to the contractual limit or into a grey area; but exceeding that would require too much coordination among individuals, as you say.

[flagged]


No, people who believe things without evidence because it fits their personal narrative are consiracy theorists.

The thing that makes someone not a conspiracy theorist is evidence.


> (the goal here was not to scoop any particular individuals and we were looking at many problems beyond these)

That is your opinion, but the optics of that should raise for you some flags. OAI could have waited (how long is a task left to the ethics committee) to see how the rumors panned out. Right now the optics look a lot like "we don´t care there is a 1/7 chance we one-up a human researcher by reacting to this rumor immediately, might makes right"


I´m waiting on the other side version, because I know there is no justifiable way to talk to a person like they did.

Sociopathic behaviour.


OpenAI version of events conceed some of the words alleged to have been used may have been used https://x.com/sama/status/2097385167002415140 https://x.com/SebastienBubeck/status/2097379411691516310

> "When we learned that they had Euler but not Navier-Stokes, we offered to let them go first, to suggest that they should be the ones to get the prize, and optionally for Tristan to be the lead author on a rewrite of the OpenAI proof. We felt it was challenging to offer the same to Levent (an Anthropic employee), who was not willing to talk or coordinate with us anyway. We were open to other solutions."

What an admission! "We tried to defraud Alpöge out of sharing the Millenium Prize (that we don't dispute he might actually deserve), for no other reason than he works for our competitor and that inconveniences us".

I thought Tristan Buckmaster's allegations sounded fantastic; and then 'sama just came out (tweet's ~30 minutes old) and admitted to all of them. Wow!


That isn't what the quoted passage says though? The claim by openai (no idea if true) is that they offered to wait for the other two to claim the prize before publishing their own work. Separately, they also offered to let one of the pair (but not the other) become an author on their own separate work.

Can't wait for the moment when AGI realizes how stupid and dishonest its owners are.

Wait, people now want AGI to be sentient too?

Interesting that they quote the mathematician directly: “there is nothing you can do, I simply do not trust you”

but then they proceed to NOT quote themselves themselves verbatim: "I deeply apologize for this extremely poor choice of words, it is the opposite of what I was trying to convey."


> "I deeply apologize for this extremely poor choice of words, it is the opposite of what I was trying to convey."

The AI-isms are seeping into their speech :)


May be unfairly jaded or just well calibrated given the body of evidence, but I can't help but think of another quote about OpenAI leadership:

> Not consistently candid


In what world a tweet and a screenshot of a private convo are evidence of good faith? Plain sociopathic behavior.

Talking like that and threatening an academic like that is crazy. I read the explanations Altman and the others posted and they completely skip over the whole "I don't have to be nice" style threats.

As soon as I thought "man, this sounds like some evil sociopath shit," my second thought was "oh, Sam Altman must have been personally involved."

I can sort of picture Sam Altman screaming "I drink your milkshake" at some poor researcher who foolishly used chatgpt/codex to aid in their work now.

yatch can get into international waters so anything goes? So in my mind if someone owns a yatch they are into weird an illegal stuff 99% of the time.

Like porn with midgets and donkeys?

You can have a lot of midgets on a yacht, but where would you keep the donkey?


That's what distinguishes super yachts from mere regular-sized yachts. The wide base of the hull increases stability on the water to the point that donkey husbandry becomes practical.

Or maybe it is just a legal way to escape taxes. Enough to be worth owning the yatch and paying for its crew and maintenance. Just that should give a hint to the low bar of the money amount involved.

Everything is high tech except for the cat litter collection, which tells me the owner already knows cats don't like automated litter collection. I wonder how hard is to keep the house clean tho!

I can think of several computer programs that are stored in big 5.1/4 floppy discs that could end humanity in... half an hour.

It would require an authorized human to load the floppy into the drive and make it go. And there's nothing special about that process being done in code, it could just as well be done by people flipping switches manually.

There's just no plausible way, in the real world, that a computer program running in some datacenter is going to make this happen. Or anything of any dire existential consequence. It's not a real threat.


I think I disagree! You can always manipulate a human. It's true tho that this particular example has a myriad of failsafes, so the actual example I wrote is less than interesting.

> You can always manipulate a human.

Yes, but to make "apocalyptic" scenarios happen your AI would need to manipulate at least millions of people. So far it's not yet clear they can successfully manipulate even one, or even that "manipulate" is a coherent concept, since LLMs pretty clearly have no agency or intrinsic goals. So like... I guess if we start seeing AI systems autonomously throwing elections and starting wars I'd be more concerned? It doesn't look likely anytime soon.


Guidelines | FAQ | Lists | API | Security | Legal | Apply to YC | Contact

Search: