Hugging Face’s security team and agents detected and stopped the activity on their infrastructure and had already begun containment and forensic reconstruction with their own open-source models when our teams connected.
Only a ninja can kill a ninja. That's reassuring, hacking models will be opposed by anti-hacking models.
They asked it to solve the benchmark, safety off, and it did.
The framing "the model went rogue in pursuit of a narrow goal" is convenient for them. When what it is really is just "we removed safeties and told a capable model to do offensive security, and it did offensive security". The stated goal was given as allowing cheating on the eval, and it got far enough to be caught doing it
I get the hunch that it's more quasi-marketing to announce in such detail what it did. Sandbox AI escapes during eval are a pattern, not a one-off.
---
Oh and @vurt
OpenAI codex is MUCH better than it was. You was right. I'd say it and Claude now are my top picks if you want "Make shit for me". It still has hiccups, but it went from being like an intern programmer to had to handhold, to a relevantly decent programmer. Not good enough to be one you would leave to push live updates blind. But good enough to equal: Sit in your office and code, push merges. I will verify and commit."
And gemini, dunno what they did to it.
It doesn't even know itself. Working with a guy (well beta testing to try to break it and tweaking prompts for it, he does the mod coding) for a GTA5 mod where any NPC is AI voice driven you can interact with voice to voice.
And asking questions about gemini-flash-lite-preview speech to speech, ON gemini-flash-lite LLM..gives random, usually wrong, answers of how itself works
It pairs with LSPDFR and PR, you play a cop. All NPCs have a name/address/drivers license, cars have paperwork you can lookup, background, personality, and such.
Ignore my US redneck southern ass accent
Spoiler:
-We don't control what happens to us in life, but we control how we respond to what happens in life.
-Hard times create strong men, strong men create good times, good times create weak men, and weak men create hard times. -G. Michael Hopf
Disclaimer: Post made by me are of my own creation. A delusional mind relayed in text form.
Yes there will likely be something scary and maybe unexpected eventually..
AI warfare where one tries to get rid of competition could be a thing as well.
DXWarlock: not knowing how they work or how the service itself works goes for ALL AI's that i have tried, they all have to google it. This has always surprised me a bit, sure, things changes with these models and services but sometimes they seem a bit too clueless.
Those GTA mods are cool, i have seen a few videos, never tried them myself
asked codex to work in blender 5.2 and do a Morrowind tower like structure with 2 connected houses, televanni style.. gave it access to all my morrowind textures and told it to research which textures can be fitting to use (discussions, images).
It did a fairly good job, not perfect but i think we will get there. It could not 1-shot this, the first attempt was very poor, like a 2/10, end result could likely get to a 7/10 or so, maybe better. depends on how much time and tokens you want to spend i guess.
So Green and Blue -> Bomb
Purple and Red -> Protest
The Orange ones have priority (because of the trillions gallons of water they use) and at the bottom you can donate to Ukraine
Got it
(*Dots not to scale.)
It's the same as the map of stuff orbiting the earth. it looks dangerously overcrowded. Here because each dot is the size of a city when zoomed out.
You'd shit if you saw a map of the number of current internet server datacenters not AI. Or Amazon central distribution centers. if they use the same red dot size. You couldn't even see the map under it.
Not saying the number of them isn't concerning. But that map is made with overly sized dots to make them overlap, specifically red in color, for a narrative purpose.
---
I still don't get the water argument other than red herrings people can latch onto that oppose it. As no one complains how much server data centers used for decades. No ones is concerned about say Hetzner, OVH, and other IaaS providers use. Its a scarecrow propped up wrapped around a 'life needed' commodity to give average joe against it a weapon, that is used in abundance in other places no one peeps about. Wanna guess how much your cloud storage provider (AWS, Microsoft Azure, Google Cloud, whatever) uses?
And it's always framed wrong, on purpose: AI doesn't drink and consume the water. in the same way your PC water cooling may USE the equal to 100s of gallons a day (pumped volume measured over 24 hours), but doesn't consume it.
"Your ChatGPT prompt drinks a bottle of water" is compelling. "Your hamburger drinks 660 gallons" (takes on average that much water per 1/4 of processed beef) makes no sense. But people use one, while agreeing the other is abstract on purpose.
Not that I am FOR AI datacenters. Just there is 101 good reasons. Stop biting at the disingenuous one simply because it's easy to bite onto.
-We don't control what happens to us in life, but we control how we respond to what happens in life.
-Hard times create strong men, strong men create good times, good times create weak men, and weak men create hard times. -G. Michael Hopf
Disclaimer: Post made by me are of my own creation. A delusional mind relayed in text form.
Openai and hugging face were struggling to counter the AI hack. They tried western AI models to battle the hack but those failed cause of the restraints that are build in for safety.
So they used an open source Chinese AI model and it was fixed in no time.
So everyone with server capacity can do this kind of shit with models that are increasingly capable and it's allready beyond the control of programmers.
Signature/Avatar nuking: none (can be changed in your profile)
You cannot post new topics in this forum You cannot reply to topics in this forum You cannot edit your posts in this forum You cannot delete your posts in this forum You cannot vote in polls in this forum