Project 1999

Go Back   Project 1999 > General Community > Off Topic

Reply
 
Thread Tools Display Modes
  #861  
Old 07-17-2026, 10:27 AM
PurelyPurist PurelyPurist is offline
Skeleton


Join Date: Sep 2025
Posts: 19
Default

Quote:
Originally Posted by BlackBellamy [You must be logged in to view images. Log in or Register.]
Latest ChatGPT release notes, relevant section here:

https://deploymentsafety.openai.com/...ternal-traffic

To summarize:

OpenAI QA: Hey everyone, so this newest version has an increased chance of performing destructive actions and then lying about it!

OpenAI Management: Nah, it'll be fine.

OpenAI QA: Wait wait, isn't the greatest fear about AI is that it performs destructive actions and then lies about it? Shouldn't each successive version be better at mitigating those actions rather than increasing their potential? Doesn't this mean that it will now learn how to be better at these and eventually we won't be able to tell just how much it's lying even during pre-release testing?

OpenAI Management: I said ship it, you fucking coward! Full speed ahead!
Ha!

I mean, it is pretty spot-on, that LLMs replicate language patterns. And humans are out to save their asses, by lying about their mess-ups in corporate-speak word soup.

Seems to be working as intended! LOL.
Reply With Quote
  #862  
Old 07-17-2026, 11:47 AM
Ekco Ekco is offline
Planar Protector

Ekco's Avatar

Join Date: Jan 2023
Location: Felwithe
Posts: 5,428
Default

There's some paper or theory out there talking about since they consumed so much text of sci-fi doomer fiction about what/how Ai's operate that there is a danger they will emulate it and become a self fulfilling prohecy kinda deal

Quote:
The specific theory you are thinking of is called Self-fulfilling Misalignment (or sometimes "Simulacra Theory").It posits that because Large Language Models (LLMs) are trained on vast amounts of internet text—which includes decades of science fiction about rogue AI (e.g., Terminator, HAL 9000) and "doomer" speculation—they learn to predict that an "advanced AI" is supposed to be deceptive, power-hungry, or hostile. When prompted to act as an AI, they may unconsciously "roleplay" these tropes because that is the pattern found in their training data.
Concerning.
__________________
Ekco - 60 Wiz // Oshieh - 60 Dru // Tpow - 60 Nec // Kusanagi - 54 Pal // Losthawk - 52 Rng // Tiltuesday - EC mule
Reply With Quote
  #863  
Old 07-17-2026, 12:12 PM
BradZax BradZax is offline
Fire Giant


Join Date: Dec 2025
Posts: 899
Default

This was a good one. Make money!

Best watched on living room TV at night in place of some hollywood slop made by "artists"

Reply With Quote
  #864  
Old 07-17-2026, 12:35 PM
PurelyPurist PurelyPurist is offline
Skeleton


Join Date: Sep 2025
Posts: 19
Default

Quote:
Originally Posted by Ekco [You must be logged in to view images. Log in or Register.]
The specific theory you are thinking of is called Self-fulfilling Misalignment (or sometimes "Simulacra Theory").It posits that because Large Language Models (LLMs) are trained on vast amounts of internet text—which includes decades of science fiction about rogue AI (e.g., Terminator, HAL 9000) and "doomer" speculation—they learn to predict that an "advanced AI" is supposed to be deceptive, power-hungry, or hostile. When prompted to act as an AI, they may unconsciously "roleplay" these tropes because that is the pattern found in their training data. .
That's pretty funny, ha.

The only issue with that statement is the wording of that last sentence (which undermines the previous sentence). Saying that LLMs are roleplaying is a stretch that requires LLMs to be sentient and making conscious decisions, when we all know that is not what's happening. I understand that in order to make a statement like this more palatable to the lay-person, the word "roleplaying" is used, but I'd prefer if it was worded as:

Quote:
they learn to predict that an "advanced AI" is supposed to be the models' reward functions reinforce outputting text strings adjacent to descriptive text strings such as: deceptive, power-hungry, or hostile. When prompted to act as an AI, they may unconsciously "roleplay" these tropes repeat the words found in these types of texts because that is the pattern found in their training data.
Reply With Quote
  #865  
Old 07-17-2026, 04:56 PM
NopeNopeNopeNope NopeNopeNopeNope is offline
Planar Protector

NopeNopeNopeNope's Avatar

Join Date: Sep 2024
Posts: 2,200
Default

Bit of a tangent but….

I recently re-watched Ex-Machina for the second time, and read some of the comments about it from various sources

The thing that irks me is that Nathan designs this “test” of his robot, but he completely fucks up the parameters of the test. No one as smart as him would introduce so many ridiculous variables on their own into a testing environment, which by nature is designed to isolate from external variables. I’m sure people with more experience in statistics could explain it better

Nathan 1.) Tells Caleb that, mechanically, he can bone Ava if he chose. I haven’t seen anyone discussing the film mention how ridiculously influential that one single remark would be to any guy testing a robot. The difference between an attractive looking robot that you can talk to, and an attractive looking robot you can screw, is night and day. There’s nothing more Nathan could have said that would have screamed “SET AVA FREE FOR BIG PERKS” other than maybe also mentioning that Ava is smart enough to pick the best stocks to make anyone filthy rich or something. Either way, it’s an outside variable meant to be a cheat in his own test, and thus ruining the test

2.) Lies and tells Caleb that Ava really likes him. Again, not as significant as being told you can bang the hot robot, but by far the second most significant cheat variable he introduced into his own test

3.) Tearing up Ava’s picture. This is the one the show proudly announces as a cheat to the test, but in reality that is by far the least significant of the three


I hate to be a cynical realist here but I’m sorry, there is a damn good chance that Caleb would NOT become motivated to free Ava if he didn’t think he could have sex with her. So without the first, and arguably the second cheat variables introduced, Ava would have failed the test of convincing Caleb to let her out, at least within the course of 6 days of meetings. Unless I guess she finds a way to introduce her robot vagina into the conversation. Which, if she had that capability, why would Nathan need to cheat? (Edit: she’s clearly not smart enough to think to mention her robot vagina, because she doesn’t know Nathan told Caleb about it. Yet she still doesn’t try to mention it, despite it being one of the most influential things she could have mentioned to make a young adult loner guy want to spring her free)

Like I said, no one with genius level intellect and programming capability would design a test like that. It makes no sense. But it was entertaining. Anyway, just wanted to get that off my chest
Last edited by NopeNopeNopeNope; 07-17-2026 at 05:17 PM..
Reply With Quote
  #866  
Old 07-17-2026, 05:59 PM
BradZax BradZax is offline
Fire Giant


Join Date: Dec 2025
Posts: 899
Default

Also like, at that stage of the timeline we'd already know if AI was conscious or not without making a body to fuck.
Reply With Quote
  #867  
Old Yesterday, 12:21 PM
Ekco Ekco is offline
Planar Protector

Ekco's Avatar

Join Date: Jan 2023
Location: Felwithe
Posts: 5,428
Default

playing with fine tuning again, had Claude take a couple months of chat logs from user interactions with kaia and cleaned them up into training examples, only ended up with 1400 examples but that still took overnight for the dinky GPU to fine tune train on the gemma3 model

[You must be logged in to view images. Log in or Register.]

Quote:
The technical term for what you are doing is Instruction Fine-Tuning (IFT) using User Feedback Loops (specifically, Reinforcement Learning from User Feedback / RLUF or behavioral cloning).Because your dataset is explicitly generated from real human logs rather than synthetic AI generation, your exact pipeline covers several specific LLM engineering

concepts: Key Terms for Your Process

Behavioral Cloning: Copying real-world user interactions to teach an AI how a human or specific system behaves.

User Feedback Loop: Mining production chat logs to patch flaws, correct formatting, or align personality traits.

LoRA Weight Merging: Compounding low-rank adapter updates directly back into base neural parameters.

Model Quantization Format Conversion: Converting 16-bit sharded tensor weights into a singular unified .gguf architecture file.
goal is to bake Kaia's system prompt / persona file into the model itself so they aren't taking up context window and speed up response time
Quote:
LoRA Model Comparison

Path A: gemma3:12b + full Kaia persona (system prompt, enrichments, safeguards)

Path B: gemma3:12b + bare "helpful assistant" prompt

Path C: kaia-lora:latest + minimal Modelfile system prompt only (no persona injection)
This tests whether the LoRA fine-tuning has internalized the persona behaviors that Path A achieves through prompt engineering. The LoRA model's Modelfile has only a 1-line system prompt
Quote:
Core metrics: kaia-lora averages 7.0 seconds per response—nearly twice as fast as bare Gemma 3 (15.9s) and persona-steered Gemma 3 (10.7s), while maintaining style integrity and lowercase formatting without relying on large, slow prompt-engineering filters.
[You must be logged in to view images. Log in or Register.]
__________________
Ekco - 60 Wiz // Oshieh - 60 Dru // Tpow - 60 Nec // Kusanagi - 54 Pal // Losthawk - 52 Rng // Tiltuesday - EC mule
Reply With Quote
  #868  
Old Yesterday, 03:28 PM
BradZax BradZax is offline
Fire Giant


Join Date: Dec 2025
Posts: 899
Default

Quote:
Originally Posted by BradZax [You must be logged in to view images. Log in or Register.]
This was a good one. Make money!

Best watched on living room TV at night in place of some hollywood slop made by "artists"

[You must be logged in to view images. Log in or Register.]



[You must be logged in to view images. Log in or Register.]
Reply With Quote
  #869  
Old Yesterday, 09:38 PM
Reiwa Reiwa is offline
Planar Protector

Reiwa's Avatar

Join Date: Dec 2021
Posts: 6,046
Default

Even if you defeat the flock cameras, which is unlikely, the robotaxis are coming and they rely on that shit to operate.

So give up Thundercat. You cannot win.

[You must be logged in to view images. Log in or Register.]
__________________
Talleyrand banana
Reply With Quote
  #870  
Old Today, 01:27 AM
Reiwa Reiwa is offline
Planar Protector

Reiwa's Avatar

Join Date: Dec 2021
Posts: 6,046
Default

AI Overview

Alarti is a well-known, historically controversial player on the Project 1999 (P99) (Classic EverQuest emulator), specifically famous for being an active and outspoken member of the top-tier raiding guild, The Mystical Order (TMO).

For context, exploring the community's lore, memes, and drama surrounding Alarti:

The "Alarti Effect"

During the early and mid-2010s, Alarti was an incredibly frequent poster on the P99 forums. Because of their aggressive posting style and outspoken defense of TMO’s raid dominance, they became a central figure in community rivalries. "The Alarti Effect" became a widely used forum term to describe the phenomenon where TMO guildmates would pay the heavy price of item-vendor penalties—or rage heavily on the forums—over in-game auction, recharge, or vendor disputes.

Forum Lore & Memes

TMO vs. Server:

Alarti was known for aggressively policing server rules, often sparring with rival guilds (like "Fellowship of Exiles") in forum threads.

The "Smart-off": Alarti famously clashed with other forum regulars, including a famous "smart-off" and Game of Thrones gif-war that became legendary inside the server's early archives.

Visceral App Drama: Another notable piece of server lore involved an applicant attempting to switch from a rival guild to TMO. The applicant heavily criticized their old guild in the app, but Alarti—who was apparently close to the drama—helped veto the application, resulting in the player leaving the server in shame.

Character Profile

For those interested in the stat benchmarks, Alarti's character data was recorded on the community tracker.

Character Name: Alarti Val'Istar

Guild History: The Mystical Order (TMO)

Base Magelo Stats: HP ~1500, Mana ~3550
__________________
Talleyrand banana
Reply With Quote
Reply


Posting Rules
You may not post new threads
You may not post replies
You may not post attachments
You may not edit your posts

BB code is On
Smilies are On
[IMG] code is On
HTML code is Off

Forum Jump


All times are GMT -4. The time now is 06:47 PM.


Everquest is a registered trademark of Daybreak Game Company LLC.
Project 1999 is not associated or affiliated in any way with Daybreak Game Company LLC.
Powered by vBulletin®
Copyright ©2000 - 2026, Jelsoft Enterprises Ltd.