Hacker Newsnew | past | comments | ask | show | jobs | submit | selectodude's commentslogin

There aren’t many primary and secondary colors. If you plan to have a more comprehensive system you’ll run out pretty quick. Chicago comes to mind. The newest one is pink and I’m not sure what they’d do if they built a new one.

You don’t. LLMs make vibe coding highly customized software that operates exactly as the user wants it super easy, which is awesome. However humans still haven’t caught up to the idea that it doesn’t need to go on GitHub because if it was useful to somebody, they’d have already vibe coded it themselves.

The only way to know if software does what is intended is to use it. IMO that's the real value of OSS: software that has been executed many times by many people in many different environments, which overall increases the trust in the correctness of the system(s).

> The only way to know if software does what is intended is to use it.

No, testing and code audits also work.

No amount of use of the software will tell you what else the software does beside what is intended (exfiltrate data, etc); testing and code audits both help with that.


Testing is not a replacement for use, but it can help exercise code to the extent that the author can predict how the software will be used. A significant source of test cases come from using code and finding defects. No amount of auditing or proactive testing replace real world use.

For every highly customized piece of software there are a thousand people who don't know they need it. This is the whole concept of software in the first place.

That's silly. Why waste time remaking what someone already built? What a bleak future

Because the hyper-tailored vibe coded browser front end that you like might be ever so slightly different from the hyper-tailored vibe coded browser front end that I want to use.

I don't want to put any effort into tailoring. I just want someone to figure out something that works. It's not like I can fix the internet itself, so what's the point?

You could search for “small lightweight browser with add blocking and minimal footprint” and find a few, evaluate them, and choose. Or you could use AI to build a “small lightweight browser with add blocking and minimal footprint.” And tweak it to be exactly what you want in maybe the same amount of time.

Okay? I didn’t sign a 10 year contract. We’re month to month and I use my own harness.

If they’re subsidizing my usage, that’s great.


You're building your livelihood/workflows on a set of inputs that you have no idea what they actually cost or how reliable they'll be when the VC cash stops flowing. If you're OK with that, do your thing but it seems a little foolish to me.

Push comes to shove, OpenAI could go out of business tomorrow and I could pick up roughly where I left off for $25k, which is the cost to serve GLM 5.3 Flash on four Nvidia GB10s. Granted, if OpenAI et al go kaput all at the same time, I could probably get a whole lot more compute for a whole lot less money.

Then just go back to what one was doing two years ago? I don't understand this argument.

If the market crashes they will be much cheaper to run actually, no? Hardware would flood the market.

That should be the outcome, yes.

In the event of a crash, the investors who put countless billions into this will be still be seeking to maximize their return. Even if it is just pennies on the dollar. Assets (including compute hardware) will be sold, just as they are also sold when any other business fails.

Or maybe a crash doesn't happen. Maybe prices rise to the moon instead and there's nothing we can do to lower them.

Or maybe (just maybe!) a crash never happens and there's never a huge price increase. Prices stay low-ish.

All of these possible outcomes suggest to me that the maximally-sane option that a user can select, today, is to burn it while it lasts. And then, if/when a crash or a massive price increase occurs, just adjust accordingly. (The rest of us will all be in that same boat, too.)


I wouldn't bet on hardware flooding the market. I bet the machines running in the data centers don't use traditional PCIe connectors and cards. Maybe somebody could pull the chips and put them on standardized PCIe cards, but that is not a given.

It happens already. These are plenty of cheap V100s on eBay, and PCIE to SXM2 adapters

External example: https://ebay.io/m/lV8UsD

Internal example: https://ebay.io/m/z1ygRU

V100s are three generations behind current and missing many of the features that modern inference benefits from, but they are the cheapest way to get a 32GB gpu.


I'm fairly sure most open weight model providers are serving them at a sustainable price - and I've used them enough to know that I could live with them if the big boys did a rug pull.

It seems silly to say we have no idea when we actually do, though. We know how much hardware costs, we know how to reliably run a webservice that hits an API hosted on a machine with a GPU, we know how to operate these things at scale outside of OpenAI and Anthropic (not Nvidia). VC money can be patient, Uber's profitable, yeah $1 Uber rides got us hooked and they're running the same playbook. Unfortunately the convenience is worth paying for, so it seems dumb to think we can control the beast or ignore it, or get everyone to agree to hold back.

Is there a world where OpenAI starts charging $2,000/month for what we previously were paying $20 for? What are we going to do? AWS could totally jack up the prices for EC2 instances as well, but we've come to rely on that as well.


huh? i use the plans because they're cheap and i get strong models, but i could go back to deepseek flash on commodity api pricing and be just fine

My M1 Pro MBP is 6 years old and continues to be the best computer I own, so if that’s Apple not trying, god help everybody else once they do.

[flagged]



What in that thread is particularly impressive or noteworthy? According to the benchmarks I've seen, M6 raster performance is actually less efficient than M5 in many scenarios.

Once I get some kind of settlement after getting beaten up by a cop my first purchase will be some RTX Pro 6000s.

Dude, I'm saying this with the best of intent. Get help.

Reddit might be leaking today.

Don't buzzkill his dreams of financing a GPU.

...actually: get help.


Dogs use the same x-ray machines that people do.

The Gulf countries paid handsomely to get Trump in there to keep the United States as the buyer of last resort once the rest of the world goes renewable.

Problem is that installing a bunch of sick sons of bitches to do your bidding can backfire when those some morons lose a war to Iran in a matter of months.


My only concern is that they'll sell out so quickly for so long that I'll have to wait for the iPhone Duo Two-o because they won't be easy to get until like May 2027.

150 tokens per second on a ternary model implies that it’s GPU bound, I’d bet a Q6 model is even faster because it’s existed longer and seen more optimization. You’d have to be insane to not run an NVFP4 quant over a ternary quant on Blackwell if they both fit.

Elon fucking Musk has literally every single piece of personal information of every person in the country. It’s so far past too late for any of this to matter.

I’m not sure how we start over but this data plus LLMs is gonna make it a full time job to keep your parents from sending every penny to a scammer.


Guidelines | FAQ | Lists | API | Security | Legal | Apply to YC | Contact

Search: