Hacker Newsnew | past | comments | ask | show | jobs | submit | 6thbit's commentslogin

I'd like a brainstorming mode, with limited side effects.

I want it to go and read stuff but also invoke some commands and generate reports and discuss on the results. Latest GPTs in codex plan mode always go.. write a plan. Shocking, i know, but thats not what i want on every turn when im in that mode.


It'd be hilarious that meta chose to pay openai. If there was such a deal, would it come to light in any public/official filings?

Depends on the size of the deal. OpenAI would have to disclose it in their S-1 if it crossed the threshold of being material information for investors.

Isn't S-1 just for IPO? Wouldn't it be in the 10-Q?

OpenAI isn’t public. It doesn’t file forms 10-Q.

No reason for it to. Software companies use products from other software companeis all the time.

meta already pays openai and anthropic for employee use despite having their own models available

  > As it turned out, the result wasn’t all that far away for humans either. A few days after I heard from Anthropic, we heard from Song He, an amplitudeologist at the Chinese Academy of Sciences in Beijing. Song’s group had already gotten the majority of the result. They’d used some AI assistance, based on GPT-6, but not the kind of one-shot almost human-less approach Anthropic used.

Please have AI come up with something no human is also about to solve?

This gets me wondering why ai labs aren't proposing their own millenium prize type challenges.


> Please have AI come up with something no human is also about to solve?

No human-only was close to Navier-Stokes. The team that was close was also using AI.


Fair point, and i think that's fine. AI labs could just not jump themselves into such problems and let humans have a go at them, whether ai-assisted or not. As in this case with a regular budget one researcher could've access to.

So human vs. AI isn't really relevant here - you just want the labs to back off.

Huh, didn't know the font could control variations over nth glyhph!

Could one make in theory a regular monospaced font to just render glyphs past 80 chars of a line as bad as possible?


There must be Pro Sailing kits already under the price of a few of these things.

Any good starting point for a fisherman not interested in stashing fish but just capturing and releasing?


They only anthropomorphize agents when convenient.

It's really a matter of iterating on the problem and validation right?

Models make progress on coding and math because they can write tests and proofs to an extent. Many industries that are more 'physical' and require performing experiments lack that instant feedback loop. Find a way to close that loop and AI begins to look useful.

But try and convince companies to invest on closing that loop just to see if the current models work well on their problems or not? Tough sell. So Anthropic just shows them, hey look, this is possible and if you don't do it I will.. so they fold.


>Models make progress on coding and math because they can write tests and proofs to an extent. Many industries that are more 'physical' and require performing experiments lack that instant feedback loop.

This is basically what they targeted with this approach. They can't automate the experiments since they are often bespoke towards certain goals or even feelings and assumptions based on sage technician knowledge that isn't really taught in any one place. Instead, they tried to automate the process of searching for candidate targets to then test in downstream lab experiments.

Seems exciting, but this sort of thing has been done for a while with just about every single ml classifier method out there for all sorts of biological data. Just yet another way to slice the pie.


The article talks about the lab, so it's not quiet at all.

Lands as an active threat. Maybe they're serious about this research or not, but for sure medical companies doing this sort of research will consider upping their AI budget and connecting their labs, etc. to avoid "falling behind".

So, the AI labs benefit either from achieving something they could market or from the peer-pressure imposed to companies in the sectors they get their nose in.


when they say "claude" what model do they mean ?

no mention of opus/mythos/fable or anything..


seems like mythos 5 did most of the leg work, they refer to it in the technical report linked at the end: https://www-cdn.anthropic.com/22573675ada52a8ca8a97a1a4b4326...

It could be an internal model as well.

Could also just be hired specialists internally and it's pure marketing (even if they discovered something new).

lol, it's true, Claude in the article is so anthropomorphized that you could read it as if it was a human just as well.

Guidelines | FAQ | Lists | API | Security | Legal | Apply to YC | Contact

Search: