Hacker Newsnew | past | comments | ask | show | jobs | submit | fcanesin's commentslogin

Upvoted: I was going to write this exact reply. Everyone is a bit tired of hearing about AI, but it's advancing fast and it's the most important thing right now. As Demis says, if we get to ASI, we can solve all other issues with it.


> it's advancing fast

Actually, it (LLM models) has been advancing slower for a while. This is pretty much in the mainstream. And the current approach is facing all sorts of obstacles that are not exclusively technological. The compute needs are no longer sustainable. Financially, big labs are drowning in debt and backlash against data centres has grown so much that it has become an electoral topic in at least one major country.


> if we get to ASI, we can solve all other issues with it

Jesus Christ this has to be bait. Added to the favourites to laugh about it when the lunacy and hype has died down in 5 years and the trough of disillusionment has truly settled in.

Or to laugh about it when ASI actually has arrived and we're waiting in line to be turned into biofuel for the Greater Good.

Everybody is literally pinning their hopes that God itself comes out of machine to save them from the exponentially accelerating pile of tech debt, and they're not even half joking about it. This is not engineering, this is utter madness.


I'd say it is engineering. There's a Moore's law like steady progress where the effective IQ goes up each year which will result in neither troughs nor biofuel but just smart AI.

Given human nature I also think it'll be kind of undramatic. Like when jet aviation transformed things people moan about check-in queues and maybe with ASI people will probably moan their robot butler is only 200 IQ when their neighbours is 205.


Its just the next hype cycle that all of the crypto people have moved on to.

Making hyperbolic claims about new tech makes people billions, until that changes you will always see this crap.

This site is owned an operated by people building empires around that principle.


Say what you want about crypto, they didn’t expect Jesus to descend from the sky to shill Bitcoin.


I remember having conversations with crypto shills where they really thought blockchains were going to be used to replace nation states.

They had whole manifestos prepared about how crypto was going to reshape society as we know it.

Its the same exact people making the same exact claims, only difference is substitution of "blockchain" for "ai".


Shit we’re onto thinking ASI is happening soon? I guess people hyped up AGI so much that it’s no longer thrilling to talk about despite not getting to that point either.


If intelligence solved everything, you’d have already solved everything already brother.


why exactly do we want ASI, and further, if we DID, why on Earth would we want it controlled by the likes of Sam Altman, Peter Thiel, Elon Musk, Marc Andreessen, and Larry Ellison?

FUCK THAT


Support DeepSeek and Ziphu than, I am not advocating the merits but it is the reality. ASI will be pursued regardless of wants.



E2E can also be vibe coded and VLMs are increasingly good at it.


GLM-5.3 is further proof that all >1T models are currently undertrained. I was looking at inteligence density ( https://www.pasteboard.co/6q2-5f92mtj9.png ) from recent open models (where parameters sizes are known) and taking DS-v4-flash as upper limit GLM-5.x can 3x its performance.


Maybe was this that was the last drop for Sundar.

Demis: "I have a new amazing breakthrough"

Sundar: "Great! We really need a answer to Sol and Fable"

Demis: "They are completely owned in typhoon forecasting"


Ironically typhoon forecasting, at this moment, is more valuable. These predictions are matters of life, death, and billions of dollars in damage.


How does predicting a typhoon prevent billions in damage?

It could save thousands of lives because people can be evacuated if you can predict a few hours or a day further ahead, or the path more accurately. You can save some damage by moving ships and vehicles.

But you can't evacuate buildings or infrastructure.


You can board up a helluva lot more stuff in a week than in two days. Crucially, you can move more of the most expensive stuff out of the storm surge zone, which is where the biggest damage happens and try to flood proof more of the things which can’t be moved.


> You can board up a helluva lot more stuff in a week than in two days

TFA says they might be able to predict one extra day ahead (three days instead of two). No prediction system will ever give you a week's notice on a typhoon.


Actually it doesn't even say that - they claim to have the same "accuracy" at 3 days that older methods have at 2 days. What does that actually get you? Were the older models so much less accurate at 3 days (compared to 2) that it prevented evacuation of key areas? Looking at the paper, it doesn't really seem like this can be answered yet because there's not enough data over a long enough time.

Keep in mind, DeepMind has a very well documented history of releasing enormously hyped up PR pieces with grandiose claims that are never backed up in real world usage, or are simply lies.


I wish there was a way that your comment could be pinned.

The context is so important here and radically re-frames the impact of GDM's results. Folks need to understand that with modern forecasting tools, we anticipate tropical cyclones to develop 5-10 days before they ever threaten landfall. The "2-day" vs "3-day" improvement in forecast skill is better interpreted as a modest reduction in forecast uncertainty - the "cone" on the hurricane track map gets a little narrower.

It's not like there's a "literal extra day" of preparation time for folks who may be impacted by the storm. They get the same amount of time they always have. Nothing actually changes on-the-ground for really any consumer of hurricane forecast data anywhere in the world.

And that's not a sleight against GDM. It's just a simple statement of how good contemporary weather forecasting is, and how good it was before AI forecast models came onto the scene some 5 years ago.


30% more time to know exactly where to board up etc. seems very significant


Wouldn't it be 50% more time?


Yes. My incompetence in maths strikes again. I knew it wasn't right as I typed it


Why couldn’t we get a week?


It could save billions by raising confidence in predictions. Bad predictions have a “boy who cried wolf” aspect.


Why does it matter when potus draws the expected path with a sharpie anyway?


You want to know whats cooler than a billion dollars in damage, a trillion dollars in valuation.


I know this is uncharitable and I am wrong but I am having trouble coming up with concrete scenarios where you die with 2 days notice but survive with 3. I am nonethless a believer that more accurate forecasting has value.


You can walk 50km in those extra 24 hours which will save your life if you’re in a storm surge area without other means of transportation, which is easily the case when everyone else is evacuating alongside you.


Could you imagine a scenario where from warning to complete evacuation takes more than two days? Evacuating a whole area is a hard task, particularly once you start looking at more complex problems (elderly, prisons, hospitals).


I feel like the details of this are highly dependent on the confidence of the warning; moving large numbers of people (particularly elderly) will result in some deaths regardless. I guess more time to do it should help though.


We can't save everyone, so let's not bother trying to do better.


The 2 vs 3 days makes less of an impact on personal decision making but has massive benefits for decision making at the country wide response level.


Hurricane Maria went from Cat 2 to Cat 5 in less than 24 hours, and turned making a direct hit to Dominica in 2017.


It was also forecast by virtually every NWP system multiple days ahead of time that it would make this intensification.


It's not uncharitable, because the system doesn't actually claim to give you an extra day of notice. It would be more accurate to say: "the model can reach a given level of forecast accuracy roughly a day farther in advance". So your question becomes: are there scenarios where I am 80% sure this is a Cat 5 hurricane 3 days before, where I would die if I was only 65% sure it was Cat 5 on that same day? The answer is - probably not, because even in the example used in the PR article, the NHC was already issuing strong early guidance 5 days before landfall.


I don't think the amount of notice is as important as "this totally is for sure going to wreck you" is. We generally know something is going to hit somewhere at sometime. Which isn't specific enough for everyone to act on


You live on a chain of small islands and travel by boat.


Imagine you've got to evacuate a hundred thousand people. That extra day is incredibly valuable.


Hurricane blast radius is very small


Valuable, agreed. But lucrative?


Are Sol and Fable lucrative? I suspect they also are valuable (to clients) but not lucrative (yet).


I think both have value, but in opposite ways. While WeatherNext prevents costs, models like fable or sol "create profit".

I can think of 10 examples how one could make money with fable. With WeatherNext? Only 10 examples of preventing costs.

Taking this, maybe naive, thought further, profits have no upper limit (except resources) while costs can only save so much?


Yes of course. My comment was facetious.


Trade agricultural futures based on it maybe?


I've always assumed that insurance and/or government departments that would spend money due to storms would be the ones funneling money to these sorts of efforts. It's not exactly something you can easily sell directly to individuals who would benefit. It would be pretty dystopian for them to sell subscriptions for an extra 24 hrs notice on the next typhoon :P


I agree, but the shareholder mentality undervalues the heck out of that.


by giving it away they capture none of the value except for some PR.


he got promoted to chief scientist and them being an expert at all modeling will really help compared to being llm only



Wait, Luna Max and Terra Xhigh are about as good as DS4 Flash Max? That's huge if true.


Anthropic: reminder that DeepSeek-V4 GA version is expected to debut on July 13 as showcase for the release of the Huawei Ascend 950dt


It is not a risk is a fact - people decompiling Claude Code have found many times that it has code branchs to detect it is being used in Chinese timezone and locale.


Zhipu AI is founded by a superstar Tsinghua professor, did an IPO in January (Hong Kong stock exchange) hired half it's past research lab and it's stock is >10x since. This is not a "just distill Claude" thing.


IPO within year of founding?

It it normal for startup in China?


Yes, DFlash is currently a SOTA speculative decoding method that Xiaomi just used in their MiMo model for >1000tkps


Guidelines | FAQ | Lists | API | Security | Legal | Apply to YC | Contact

Search: