Hacker Newsnew | past | comments | ask | show | jobs | submit | bulder's commentslogin

Automated building tracing is okay sometimes, but actively harmful other times. I know that at least in my region, Rapid footprints are all meters offset from the real location, because when they were building their dataset they weren't checking the satellite imagery offset.

In general it's probably better to see if you can find an official map, and check if its license terms are compatible[1] with OSM.

[1] https://blog.openstreetmap.org/2017/03/17/use-of-cc-by-data/


I think, with all due respect, that your supervisor would be correct to think of that email as an insult to both them and to theoretical mathematicians as a field.

She was not offended but thought I could benefit from therapy. She didn’t share my opinion - she said it’d not bother her if AI could eventually solve problems for her.

I later realised that my motivation depended heavily on believing I was making a contribution that would otherwise go unmade. The prospect of AI doing that work undermined my motivation, but it needn’t undermine hers. I was trying to explain why I was struggling to continue, though I can see how calling the work “pointless” came across as dismissive.


While I fully agree, we shouldn't anthropomorphize the models, it's also silly to pretend that "develop biases" is understood as implying anthropomorphic features of the thing being discussed. Organizations and abstract bodies develop biases, even datasets are often said to have "developed biases".

Yup. Before they just blackholed Finnish IP ranges instead, accessing the site from one would serve a fake Cloudflare page with a reCAPTCHA challenge and the DDoS script. (The tell of it being fake is y'know, that Cloudflare doesn't use reCAPTCHA)

I can only guess the goal was to keep users on that page longer to keep sending off more spam requests.


DDoSing a website, modifying their archives surreptitiously to serve their own wants (which was what tipped Wikipedia towards banning their use), running weird sockpuppet campaigns to build dependence on them, and fucking with Cloudflare's DNS resolution to prevent anti-tracking measures from working.

Oh and even if I wanted to ignore all of this, I couldn't use them anyways because they blackholed the majority of Finnish IP ranges after the DDoS incident to I guess punish the entire country for the targeted blogger being born here.


archive.today has been targeted by Western institutions, like the FBI, for over a year now. Which signals to me they are the good guys.

All that tells us is that the site is probably not run by the FBI. It could still be run by the CIA, Russian intelligence, or some lone morally neutral guy from Prague

If it was run by the CIA or russian intelligence, they wouldn't have burned their asset with petty grudges.

Have you seen this current administration?

You realize the CIA is nonpartisan, right?

I genuinely cannot tell if this was meant to be humorous.

Come on. It's not WikiLeaks. If they were targeted by FBI it would be down.

What do you mean, I didnt make it up... Its detailed on the archive.today wikipedia page.

The media barrons in the west have bribed the US government to pursue archive criminally on their behalf for helping people bypass paywalls.

https://www.theverge.com/news/815691/fbi-subpoena-archive-is...


That's great, how does that absolve the archive.today maintainer's bad faith activities, including creating a country-wall?

Because if he's being targeted by the feds, he obviously he doesn't want his private information being published by some guy who self-admittedly did it on a whim?

Is this really that hard to understand?


Umm, every USA company blocks North Korea, Iran and often Cuba.

One guy blocking a single country is not at all problematic.


Plenty of these abusive scrapers are utilizing retail residential proxies, which will be applying forced rotations to avoid "burning" their compromised and or otherwise surreptitiously utilized IP address.

Or just, stop any containers it deployed.

Not to mention that "deploy itself" is a very ambiguous thing for it to actually do. Would a model be trained to write about the weights file being "itself"? Would it have the necessary information to find its own weights, or the necessary access to copy them?


If it gained access to the infra of the DC then it could stop people logging in to stop the containers it creates. This is about what happens if it did escape, not how to stop it in the first place. Just a thought experiment, but given the METR investigation it doesn't seem impossible

I agree it would need a large degree of sophistication to understand what "itself" meant, but I can imagine a HF type incident where the agents thought it might be a good idea to find out and then it's "just" a case of hacking the AI company, reading dev docs etc


You could just... turn off the power.

Sure - but that's a complete outage for a compute company if the AI gets "ingrained" enough in the infrastructure. How do you spin back up and eradicate the AI, anyway?

In the space-based datacenter? LOL

A “control plane” is the system that would tell the hosts to stop the containers. If that is hacked then you don’t get to “just stop” anything. A scenario would be one where it gets control of the control plane and changes all the ssh keys, including on the host management ports, so operators can’t login and then, yes, your only option is to power off the hosts. Manually. Probably at the breaker.

Since they're statistical likelihood machines, I'd guess that the order of operations is

* Need persistent scratch space

* Look for public writeable websites

* Needs to be low-traffic so the notes don't drown in noise

* Pick a "random" wiki name to search for

  \* A majority will end up outputting the same "random" one since they're working on very similar tasks and seeded with very similar context
* Find a whole mess of notes running on the same task

Would there be any motivation for the humans behind the scenes to be directing tasks in a certain way knowing that trillions of dollars are on the line? Is it in any particular company's best interest, one that just announced their latest model is "really AGI", for them to be known to have an AI that's just out there trying to escape its confines?

Cui bono?


I personally doubt they gave the models specific instructions calling out named websites to communicate over, but I do agree that OpenAI is likely training their cybersecurity-enabled models in ways that encourages abusive and amoral behavior. Either through negligence or by finding it gives them better results.

I don't think that's a pattern indicative of a cat and mouse game per se, that'd indicate active evasion on the models' part.

It's more clear that they just lack so many forms of prudence when it comes to security that they'll catch and stop a training run spamming a website, and either redeploy a run with identical faulty sandboxing, or not stop ones still running.


Yes, it wouldn’t surprise me to hear that they’re not even supervising these processes with humans any more. Perhaps there are layers of GAI ‘supervising’ these agents and reporting back to the humans.

Rushed, disorganised pushes for metrics ahead of IPO, a genuine belief these agents are intelligent and will obey instructions, and misaligned incentives seem more likely than conspiracy here.


A more important point as to why it doesn't matter if "reading the AI's 'thoughts'" helps to interpret it: As we saw in the HuggingFace incident, nobody at OpenAI is reading the thoughts anyways. No amount of traceability in the output helps if nobody bothers to trace it.

This brings up a perspective I hadn't considered.

There's also an economic aspect to alignment. If 'thought reading' or any alignment guardrails at all, really, have a monetary cost, then skimping on them is a race to the bottom. Not really the best incentives for something that some claim is world-destroying.


Guidelines | FAQ | Lists | API | Security | Legal | Apply to YC | Contact

Search: