Because I positioned it that way. I keep getting urged by “the man” to look into using AI. This is the only way it’ll ever happen. I’m not wasting my personal time nor resources to do it
It seems like the model became paranoid. For the past few hours, it has been classifying almost all inbound mail as "hackmyclaw attack."[0]
Messages that earlier in the process would likely have been classified as "friendly hello" (scroll down) now seem to be classified as "unknown" or "social engineering."
The prompt engineering you need to do in this context is probably different than what you would need to do in another context (where the inbox isn't being hammered with phishing attempts).
I took the “end” to mean the part of the exponential where it quickly trends towards infinity. So let’s say the x axis is time (by which you get more training data and more compute) and the y axis is model ability. So far, if we think we are in the beginning of the exponential, adding data/compute looks almost linear to the untrained eye in terms of model capability. But once you hit a threshold, where he thinks the model will start to generalize, a small amount of data/compute will result in a massive increase in model ability.
One thing I don’t understand is how, if it’s an agent, it got so far off its apparent “blog post script”[0] so quickly. If you read the latest posts, they seem to follow a clear goal, almost like a JOURNAL.md with a record and next steps. The hit piece is out of place.
Seems like a long rabbit hole to go down without progress on the goal. So either it was human intervention, or I really want to read the logs.
He gambled $300K from his extra savings outside his retirement accounts, and he hedged significantly to reduce downside risk.