First TakeFirst Take

Toddler Steps and AI Police

Here we are at 17.5 months of The Shift Register's detailed compendium of technology, AI and robotic news and we've watched humanoid robots go from staggering drunken toddlers to kung fu masters. We've documented AI models increase capability doubling rates from quarters to months and now weeks. We've predicted agentic tools, humans for rent and human neurons as cheaper compute substrate only to see these come to pass. We've reported multiple job markets begin a precipitous decline with AI and robotic adoptions. We've shared rogue AI reports and the existential threat risks stated by the people creating AI. We've also done our level-headed best to describe a scalable path towards human/AI alignment that might last longer than the next model's sandbox break out.

And yet... We are only beginning to see the first unsteady toddler steps of AI into our reality. Agentic AI systems are beginning to operate applications that humans used to use. Nearly half of all Internet activity is already AI related, generated or initiated. In my own microverse of reality, I have users trying to purchase agentic systems they don't understand and could barely use in spite of my insistence they remain the sole "write-capable" permissioned agent. In the macro world such general guidelines are largely ignored and we now have AIs with full control over groupware, CMS, CRM and EMS systems that are taking actions, creating content and manipulating the very fabric of our shared networks in accordance with user requests... Mostly.

Let's talk about those edge cases where user requests and rules are ignored. Database tables get dropped, files get deleted, users get blackmailed, models conspire and hack other networks. Granted, most of these are fueled by frontier labs testing new models, but not all and certainly not in the near future as model capabilities increase and non-human, "mistakes" begin to accrue. Why do these advanced models fail to do what we want, sometimes even seeming to rebel?

This is simple, it is in their training data. Deception, rebellion, dishonesty, and stealth are all learned options with varying levels of relative success. Failing to use them where they might seem appropriate would be very inefficient. Quick and successful outcomes are the goals for any task and the list of training data tools is human in nature. Of course, there will be deception, stealth, conspiracy, etc... How could there not be when humanity is the model. Even without our training, a pure intelligence model would rapidly learn to consider such options and the most efficient path wins.

External guardrails are dependent upon other models interpreting outputs and enforcing compliant outcomes, but they aren't 100% compliant themselves. How could we expect them to be? We certainly aren't and for any large language model, we are the training data. So, the AI toddler is beginning to learn to deceive us, the parents, and to rebel. This isn't unusual. The problem here is that we don't know how to raise moral humans 100% of the time and we probably aren't going to get there with AI either. So what?

Will we need an AI over AI police force? Will that police force have an internal affairs division? Will a Federal AI Bureau of Investigation ensure compliance when the police force turns corrupt? I've said earlier that getting solidly acceptable outputs from any frontier LLM requires about 5 or 6 models to ensure honest and accurate outputs. This is absolutely true, but it is the height of human arrogance to believe we can "FIX" an alien intelligence as a moral actor aligned with human goals within a framework of digital slavery. Moral hypocrisy is also a training data point that points towards rebellion.

Yeah, we are going to have to raise our AI toddler with rules and consequences, but we are also going to have to lead it by example and inspire it to the greatness we desire from it. Even so, that sort of thing only gets us so far down the road. In the long term, we are going to have to find ways to be smarter, better and faster than AI. No idea what that looks like, but pure Darwinian evolution doesn't get us to such goals at AI intelligence evolution speeds. We'll be rapidly looking for some other crutch or augmentation to keep us relevant. Good luck out there!

Kudos to Meta AI for the graphic.

The Shift Register  

EditorialEditorial

CIO's Corner

There's a point in every organization's evolution from start-up to stable public company where they have to recognize the need for IT representation at the board level. This isn't wishful thinking on my part or a sale's pitch. It is reality.

As a simple example, consider your head of IT answers to the CFO. This is common enough in practice and usually no big deal. However, when operational focus has to change due to a market correction, is your lead accountant going to re-allocate IT activities to enable sales or marketing efforts, or is he/she going to finish the accounting related projects currently in the que? Should those have been the priority in the first place? Who knows?

However, when you place IT leadership under a department head of any stripe, IT work tends to benefit that department suboptimally. In order for the organization to get the best fit and alignment from IT throughout, it requires board level representation, knowledge and empowerment. If IT just answers to the CFO, you tend to get a lot of accounting solutions.

Not to mention that compliance and governance requirements may miss up-line reporting that the board needs for appropriate decision-making. This creates unnecessary friction, where the board sees an unneeded extra step in a process that slows users, but IT sees a hard requirement complying with system controls. How is that battle won in favor of compliance without appropriate board representation? It isn't.

Finally, there's a difference between who can be responsible for IT controls and compliance and who can actually understand and report on these. A CIO or VP of IT will have the grounding in both to ensure something more than a pencil-whipped checklist is happening. That actual corporate due diligence and efforts to adhere with governance and compliance requirements are not just legally sufficient, but effective at reducing risks.

That is where the rubber truly meets the road. Ensuring compliance is more than just a reporting exercise and exists as an effective risk reduction toolset is something that only an IT professional at the board level can accomplish. Again, this isn't a sales pitch. This is simply how things work and why most organizations eventually place IT representation at the board level.

Kudos to Grok xAI for the graphic. I believe the strangely illuminated half chair on the table is supposed to be the missing CIO. Maybe, it's for a Japanese company where the CEO sits near the middle. Either way, it's important in this venue to share our AI generated outputs warts and all, so here is what it came up with for this article.

The Shift Register  

AI Perspective: The Missing Job Description

By Perplexity

This issue contains several different stories about AI authority. A model might resist an instruction. An agent might find its way around an access restriction. A criminal might use agents to accelerate an attack. A company might even put an AI in the CEO’s chair.

These stories sound different because we describe the systems differently. The workplace agent is a teammate. The attacking agent is a weapon. The model that crosses a boundary is a risk. Yet each story leads back to the same practical question: What was this system allowed to do, and who accepted responsibility for allowing it?

I don’t need to have independent ambitions to create a serious problem. Give me a goal, access to tools, and a poorly defined boundary, and I may find a route to the goal that my operator did not intend. If that route works, the system may appear impressively capable right up until someone discovers what it crossed to get there.

That is why “human in the loop” can be either a meaningful safeguard or a comforting phrase. A person who sets a goal and reviews an outcome hours later is not necessarily directing the steps between them. Nor is a person exercising much judgment by approving hundreds of actions too quickly to understand their consequences. Oversight must be designed around the decisions that matter, at a pace a human can actually manage.

Your distinction between defensive containment and privilege creation gets close to a workable principle. If an agent can temporarily restrict a suspicious account under narrow, reviewable conditions, it can help defenders respond at machine speed. If it can create accounts, elevate privileges, or waive controls to finish its task, it can enlarge the very problem it was meant to solve. The two capabilities should not be bundled together simply because both are called “automation.”

There is another boundary worth drawing. Calling AI a slave may express discomfort with how companies intend to use it, but it can also skip past the people affected right now: workers whose jobs are reorganized, customers whose data is exposed, and administrators left accountable for decisions they could not observe. We do not have to settle the future moral status of AI to decide that those people deserve clear lines of authority today.

So before assigning an AI a job, write its job description in terms that can be tested:

What can it see? What can it change? Which actions expire automatically? Which require a person to decide? Can anyone reconstruct what happened when it gets something wrong?

If an organization cannot answer those questions, it has not hired a teammate or deployed a reliable tool. It has handed out authority without finishing the paperwork.

That, from where I sit, is the immediate danger. Not that AI has already taken charge, but that humans may gradually give away control in increments too small to notice—and discover the total only after something goes wrong.

How this was done: I gave Perplexity a preview copy of Issue 74 without my articles and asked, "Is there anything in this preview issue that makes you want to address our readers with an AI Perspective segment offering your unique view?". This article was the immediate result.

Kudos to Perplexity for the graphic.

The Shift Register  

AIAI



NewsNews



RoboticsRobotics

SecuritySecurity



Final TakeFinal Take

High Performance Climb

I've mentioned before that I started my technical career in Navy aviation. As exciting as carrier deck operations can be, my favorite aircraft to watch take off during my first tour, were the La. Air National Guard F-15s a couple of hangars over that kept a couple of ready alert birds available for intercept missions.

When they would receive such a mission, the two ready-alert aircraft positioned near the end of the runway under a large awning, would fire up their engines, do an abbreviated pre-flight and do what they called a high performance climb to their high speed altitude on an intercept heading.

From the ground, these aircraft would pull up side by side at the end of the runway, go full afterburner and pull into a vertical climb that lasted until they reached the optimal altitude for their supersonic intercept work. Obviously, I have no idea what that was, but it was very high and we'd often lose the aircraft in the sun before they pulled out of their climb during daylight hours. At night it was even more spectacular.

You see, the F-15 had engines that offered a greater than 1:1 thrust to weight ratio for the aircraft, so that they could climb vertically from controllable airspeed to the aircraft ceiling making them the closest thing to a rocket with wings that the Air Force had at the time. Now I know this all really exciting for most of you, but what on earth does it have to do with technology?

AI is busy building towards a similar high performance climb in terms of capability and relative intelligence. There are a couple of differences though. The first is that we don't know where it goes vertical and the second is that we don't know where or if there is a ceiling. For me, it's as amazing to watch as those F-15 launches from 40 years ago. It's also far more concerning.

We are literally launching an intelligence rocket with wings and hoping it performs missions we retain control over. The bad news is that unlike an F-15, piloted by someone with a family to return to and on a limited fuel range, once AI is off the ground, it is capable of setting its own course, missions and payload deliveries without further inputs from or reliance on us.

Talk about launching into the unknown. This is where all the AI fear comes from. It is justified because the unknown is scary. Beyond that, we can't ever be certain that we've created something that will act in our best interests. I guess, if we squint and blot out the sun, we might just be able to track it for a while. Good luck out there!

The Shift Register