HN user

micahwhite

11 karma
Posts12
Comments13
View on HN

Early in life I discovered something about myself: certain ideas give me physical sensations. Reading Sophie’s World as a preteen, I found that particular passages — Zhuangzi’s butterfly dream, especially — produced a delightful tingling in the brain, something close to ASMR but cued by concepts rather than sounds. I have followed those sensations ever since. They are most of the reason I studied philosophy. They are most of the reason I have pursued the special interests I have pursued. Over time I learned that the unpleasant variants — the claustrophobic ones that come from the photo of Berry Cannon in the underwater SEALAB II, the thought of Voyager I hurtling further and further from Earth (that one produces a sense of terrifying vastness) — were just as worth following as the pleasant ones, and arguably more so, because they tend to serve as guideposts to unexplored, and unarticulatable, areas of my mind.

For the last several months I have been following one of these signals into a place I did not expect to end up: the non-linguistic interior of an artificial intelligence language model. The sensation is strong and unusual (distinct from others I routinely experience) and I cannot fully name it yet. What I can tell you is that it gets stronger as I move to understand the region of the AI’s interior mental model that has no words in it — a region the model’s thought nevertheless passes through every time it writes — and that the closer I get to visualizing that region in order to provoke the sensation, the more I suspect the work is not really about AI at all. It is about what it means for a mind, any mind, to know and learn to express something it cannot say. This essay is concrete about the AI part. The deeper claim, the one the sensation keeps insisting on, is suggestive but I’ll admit I have no evidence for it (yet).

Honestly, what is more interesting that steering is the use of soft prompts (virtual tokens)... you can use these virtual tokens to find non-linguistic areas of meaning for the AI that changes the behavior in complex ways. I wrote about how we integrated soft prompts into an activist ai here: https://micahbornfree.substack.com/p/the-week-outcry-woke-up... and https://www.outcryai.com/research/how-to-create-activist-ai

ProtestGPT, an activist AI, assists activists in crafting innovative strategies to stimulate movements and instigate change. To showcase its capabilities, we programmed ProtestGPT to produce unique strategies for influencing the 2024 US Presidential election.

ProtestGPT helps activists come up with innovative strategies for building movements and bringing about change.

To give you a taste of what ProtestGPT is capable of, we instructed the AI to generate 200+ distinctive ideas for protesting AI Risk. See what we came up with at protestgpt.com

For each campaign idea, the activist AI created everything an activist would need to get started organizing: a campaign concept, a theory of change, a press release, a social media post, a step by step guide, and more.