Hacker Newsnew | past | comments | ask | show | jobs | submit | dist-epoch's commentslogin

`codex -p` is the subscription equivalent. or ACP if you want to be fancy

yep. I'm saying that the managed agent API described in the article is API only.

You can give the agent a tool (or bash script) which waits for events. Agent calls it and the tool sleeps until an event happens then returns it to the agent.

Nobody wrote 13 mil lines proofs before.

I'm pretty sure you can make Lean at least 10 times faster if you unleash the agents on it.

Somebody ported Doom to run entirely in the TypeScript TYPES (not code). It took 12 days to compile.

https://www.tomshardware.com/video-games/porting-doom-to-typ...


Before transferring to life, it doesn't even transfer to a game like Poker.

You could consider ZFS a VCS, and it can easily store multiple versions (snapshots) of Visual Studio.

It's not impossible, there just isn't that much demand for it.


One has nothing to do with the other.

It was long predicted that math and software developments would be the first domain where AI was going to do major damage.

If OpenAI and Anthropic didn't get into math result dick measuring, Internet anons would have in their place, 6 months later when it got cheaper.


I think you’re missing an important distinction. “Major damage” to the talent pipeline because models become capable of original end-to-end mathematics is what the community has been discussing. But if the models rely on sniping nearly complete work then this damage is antisocial without a lot of upside, it would be destroying a talent pipeline that would still necessary for continued progress.

Which is it? I don’t think OpenAI is being transparent enough for us to really understand whether these results would have been possible without relying on unpublished information from the solution strategies of the experts


There have been about 6-8 major math breakthroughs claimed by AI. Only for 2 of them there are public accusations about the training data.

Only? That doesn't look small to me.

That we know of

It's all marketing. Don't fall for it, it's the age old strategy "our product is extremely dangerous, so fear us".

Like the tobacco companies saying all the time "our cigarettes are so dangerous, they cause cancer".

Or like Purdue Pharma saying "do not use fentanyl, it's so dangerous, it kills thousands of people every year".

They try to generate fear in their products, to shock investors into buying their stock.


I enjoy HN's brand of cynicism as much as the next guy, but this is ridiculous. Tobacco companies say that because governments require them to. They are not putting photos of diseased lungs on their cartons to "shock investors into buying their stock", whatever the hell that even means.

What even is the marketing strategy though? Generate a buzz headline so people engage and speak about the company? Isn't commenting falling for it?

Posts here with no activity can't climb gravity without engagement.


“We have built this thing that is so powerful that it will wipe out all jobs ever created. Therefore we are the only thing worth investing in.”

OpenAI says they just pointed the AI at the problem:

> Regarding the level of human involvement on our end: although a group of people was involved in our efforts, we collectively had no research-level expertise in fluid dynamics and the Navier-Stokes problem, and therefore were unable to meaningfully contribute to the mathematical content.

https://x.com/SebastienBubeck/status/2097379415747342689


They needed the general direction and idea it might work from a few folks tinkering for a year with the same models that we all have access to. The majority of credit goes to them. The idea of where to look with AI is key. Models didn't do that by themselves.

Trust: https://x.com/__alpoge__

"since on my side things were mostly me and claude having a good time yoloing random stuff in the corner rather than anything institutional"

Not: https://x.com/SebastienBubeck

"If you don't want me to be nice, then I don't have to be nice." [when clearly there's not enough deference paid to where this line of research originated]


Also: https://cims.nyu.edu/~tristanb/statement.pdf

I had planned to say on announcing our work that the results are not the important thing. Rather the important thing is instead the significance that a mathematician and an LLM model can now do all this work in a month


This is the plot of 3 Body Problem, the Dark Forest. You need to hide yourself (the problem you are working on) or the super advanced aliens will obliterate you (start working on your problem) the moment they know you exist (rumors the problem is amendable to LLMs).

Some of the involved people are dramatically naive if they believe they can simultaneously hide themselves from OpenAI while sending their arguments to a cloud service owned by OpenAI.

While I support their argument - push for stronger data and privacy protections from OpenAI and similar - it is naive to believe we can have privacy while sending our data to third parties. It's clearly better to be safe than to be sorry here. Well, clearly better in terms of privacy. In terms of the maths gold rush, who can say what's better, that probably favours those taking more risk.


Guidelines | FAQ | Lists | API | Security | Legal | Apply to YC | Contact

Search: