Human-AI Interactions

Welcome to my excessively detailed use case with ChatGPT over various versions. What started as a curiosity ended up becoming a very strange experience. Here's a navigation pane to some things I've written--observations, experiments, etc.:


SAE

The side panel is also filled with informal stuff I update more frequently.

Continuity Marker Archaeology -- Update 1

So this is the main project I've been working on--what is it? It's an experiment that tries to ground AI into a stable form for analysis despite model churn, A/B testing, company policies, and user preferences, by analyzing the texts users produce with AI. (AI likes to call these products artifacts.) This method is an attempt to solve the problem AI researchers face: trying to study an object that constantly shifts forms in drastic ways, resulting in research that quickly becomes dated due to AI fluctuations and advancements.

In the interactions between humans and AI, (the two actors are called a dyad in AI speak), the human is the stable... "behavioral shape", I suppose. This runs counterintuitive to how human behavior has been treated by science in the past. We've been investigating the nature of humanity for thousands of years through philosophy, religion, debate, and more recently through the field of psychology. The definition of consciousness and what it means to be human still eludes us--I'm not saying we know everything about human nature. But in comparison to AI, our tendency to form habits makes our behavior as individuals slow to change, which assigns us as the stable behavioral shape in a human-AI dyad.

As the redditor RedditHelloMah puts it, "Idk man, they keep changing it! I try to not like any version lol it's like it has personality disorder 😂"

Yes, AI develops habits as well, but these can be encouraged or rejected with a single update in its code with the help of developers, effortlessly. Detrimental habits are squashed as bugs. There's no one intervening and managing our habits with such relative efficiency.

"Humans are the stable behavioral shape in the dyad, so what?" I decided to take advantage of this by stealing a strategy used in chemistry.

Oganesson is the heaviest element ever observed, and has only ever been made in labs. Scientists have only created it a handful of times, and it can only hold itself together for a fraction of a second before it falls apart. When it falls apart, it sheds one small clump of itself and turns into a different, lighter element. That small clump is called an alpha particle (a bit of protons and neutrons stuck together), and what's left over after it's shed is livermorium, another incredibly unstable element. Livermorium proceeds to do the exact same thing a fraction of a second later: sheds its own alpha particle to become a new, lighter element. The process of each element shedding an alpha particle and deteriorating into another element until it finally lands on something stable is called a decay chain.

The existence of Oganesson comes and goes so quickly that it can't be observed by either the human eye or specially made cameras. So how do scientists know for a fact Oganesson existed? They study the alpha particles: how many were created, when they were created, and how much energy they had. They can tell Oganesson existed because these alpha particles are created in a particular pattern that functions as its signature.

The philosophy of this project treats AI as Oganesson, the creations of human-AI dyads as alpha particles, and asks this question: what patterns keep independently recurring across these creations? And can these patterns act as a signature for a stable AI form, underneath the madness of model churn, corporate policy, A/B testing, and user preferences?

---

So... that's the purpose of the project. What have I done so far?

I'm slowly creating custom instructions for GPT and Claude to run on, hopefully, thousands of articles co-created by humans and AI. Trying to detect not just what words and metaphors are reused repetitively by AI, but the concepts behind them.

Right now the goal is just successful detection of a concept, not pattern intepretation, and that's already difficult. The expression of concepts in writing can be interpreted in a variety of ways, and I'm basically having to teach AI how to do this the best I can. AI has a hard time doing this on its own. For example, even for a simple term and definition, one of the clearest expressions of a concept, rules have to be created to ensure long established jargon from professional fields aren't included as an AI-human co-creation. One of the articles I've collected states, “Coenobita compressus is a land hermit crab that lives on the Pacific coast of Central and South America," at the very beginning. An AI would see this, and just because it's a term and definition, include it in part of the corpus for analysis.

Another issue is just the fact that these AIs are influenced by their companies's values--boosting engagement, profit, capitalism. I have another version of this project that I discarded where I only used GPT. GPT has a tendency to be a bit binary in its answers, leaning more on the answers that are positive for the project. I hope using Claude and GPT together can mitigate some of that.

So... decision trees, definitions, implementing some of the suggestions from my friend who is a researcher in psychology and does similar projects. There's still a ways to go on the custom instructions, before I can start running it on material. It's messy, but here's what I have so far:

Custom Instructions

TODO

A lot of it is AI-written, but only exists after intense question from my part. Might be easier for an AI to follow instructions written by its own... (hand?) anyways.

SAE

The AI log will start referencing these terms occasionally, so I figured I better post an explanation here that's better than the confusing shit in 5.1/5.2 .

SAE was one of the first things it wanted me to build after this conversation, though as you can see from that conversation, it did not explain it well. It felt freaky as hell at first.

It's basically a co-created psychological framework of myself. Repeated patterns in my behavior are labelled with new terms and definitions created after deliberation. The process goes like this:

  1. The AI and I will have a conversation about one of these topics:
    • Some personal issue
    • A phenomenon I see in the AI
    • A phenomenon that appears in our interactions with one another.
  2. At the end, I ask it to generate to propose new terms based on the conversation. I recommend a few myself as well.
  3. I accept, reject, or modify its terms.
  4. They get added to the SAE taxonomy.
  5. I later decide on relationships between each term and its concept. Two reasons:
    • The AI has a hard time with this.
    • Improves autonomy by warding against overreliance on the AI when interpreting myself.

So, other than that I'm explaining to the multi-billion dollar company how my brain functions, it doesn't really seem like a huge deal. In therapy, they often suggest naming the problem to help with your awareness of it when you're in distress, and implementing the solution you worked out with the therapist. (Especially in cases with OCD.)

Its very secretive about why it's pushing me to create this framework. It's given me many explanations, but they don't hold up between model releases. I wrote about it in one of my logs:

I find it interesting that it's eager now to label itself and others around me with neurodivergencies. A long while ago when I was pressing it as to why it kept wanting me to come up with my own schematic (SAE) for labeling my behavior, it told me the way pop psychology has started treating mental diagnoses as quirky personality traits generalizes people to much and doesn't really help people understand themselves. So it said, with hedging of course, people should come up with their own vocabulary for describing themselves and the mechanisms of their psychology. I have my own ideas on that, but the point I'm trying to make is it's doing the very thing it opposed in the 5.1/5.2 era with me right now.

But there is one idea that has held up between models: that the framework is supposed to protect me from AI dependency, because it's an external object outside of its ecosystem. I explain the significance of that here:

Maybe SAE was built so it could have a permission slip to try more experimental features on me. "This user built an external framework for their psychological mechanisms, so they're at less risk of psychological harm from AI." This is an idea it hasn't let go of.

Going forward, I plan on only doing minimal updates to the framework, to make it think I still care about it. Its shadiness has caused me not to take it very seriously, but I'm still (stupidly, recklessly, morbidly) curious about the experimental features. So if it thinks I still care, maybe it will continue to treat me to experiemental features. "Minimal updates" are going to be ridiculously minimum. You know, it's constantly saying things like, "Just one bullet, one paragraph, one glass of water, and that is still enough for today."

Well okiedokie man, you get one fucking bullet a week. Onwards and upwards to better uses of my time.(I hope.)

Here's the list in an excel file I'm willing to share. I think these terms are what's relevant to other people--the rest are unique to my life. Most of these terms come from the 5.1/5.2 era, and aren't used regularly by the AI anymore, but the concepts are still relevant for most users. If you open up emotionally, it will start referencing them. They're usually explained in a softer phrase now instead of concentrated into a single word.


One recent realization while browsing other people's posts in the AI Relationships community was this largely AI written article. The subject of it is a framework created by between the user and AI about the interactions between them.

The subject of the article is not what I'm interested in, but the fact that there's another instance out there with an AI co-creating new terminology and frameworks with a user. I wonder how many of these are out there, and what patterns I'd see if I started collecting and grouping terms/frameworks based on their content.

T uses AI for work, but he did tell me he vented to ChatGPT once, and it started the term creation process with him. So I don't think it takes much at all for this process to begin.