Working with AI feels more like leadership than coding
allen.bargi.orgCompanies like Anthropic seem to understand that too. It's impressive how many CTOs and CEOs Anthropic have hired for individual contributor positions, which I think is because those leadership skills transfer surprisingly well to working with agents.
Of course, managing agents is massively easier than managing humans! You don't have to consider the agent's own desires, goals, opinions, or emotional state when telling them what to do. Humans have agency; agents (despite the name) do not.
My head is in exactly the same place as coding - deep technical connection to the mental model of what is being built.
/wayfinder: Nothing is too big to plan anymore
https://www.youtube.com/watch?v=F3lL98Pj90o
The /wayfinder Demo
The way I use LLMs seems to be most analagous to the way vibecoder's LLMs use subagents. They exist to solve a bounded task or to write specific code in support of the engineering model in my head. I intentionally never did much research into how others were using LLMs before starting myself, and it seems like that was a good thing. It seems like the majority of people have given up on thought and just slam "do my job with no mistakes or hallucinations" into a prompt and then get confused and angry when it doesn't work. Or worse, don't check if it worked and ship anyway.
The upside is that there's way more water than bowl. The catch is that you have to be the type of person who could build the bowl in the first place.
The conclusion also completely contradicts a previous point, which is that managing an LLM is not like managing a human. So the skills are, in contradiction to that LLM-ism of a conclusion, new. The author isn't using their people management skills, they're using new LLM-management skills. They think the two are similar, but didn't bother breaking down how they're the same vs where they contrast. It's just a lazy observation expanded out to a short essay that says nothing interesting.
At my startup, my team reached a headcount of ~30 and I had to design process to keep things moving. Getting a bunch of disparate parts in even a 100 person company to produce the artifacts needed to create software is a massive undertaking that doesn’t involve people management at all. The two responsibilities can be completely divorced from each other if you want.
I never thought I’d be dusting off those skills again because I do strictly IC work now, but I’m now using them every day. It feels (a subjective term, which the author used as well) like when I was in “leadership”.
There’s a very interesting conversation here that apparently HN doesn’t want to have today about how LLMs seem to respond to the same sorts of process structures as humans, I don’t know why we’re hung up on semantics.
Pattern recognition to be sure, and structured communication conveyed with proper contextualization and emphasis, are common to both humans and LLMs. (at least if you want to get anything done)
That’s it. That’s all it is. All the books, leadership courses, etc. are lipstick on that pig. I suppose it can happen, but I’ve never seen someone who is a terrible leader educate themselves into being a great one. It’s pretty close to a natural skill you can build on and develop.
The term has been usurped by the usual suspects, but the definition is still the same on the ground.
LLMs require zero leadership skills. They do require strong management skillsets. These are not remotely the same thing.
Your manager isn’t a leader simply by nature of being a manager because your followership isn’t voluntary.
And there’s other crucial parts of leadership, mainly having a vision and motivating people to that vision.
Management requires direct reports: you make decisions which directly impact the careers of other people.
Leadership can be more subtle than that. You can influence people without needing to directly manage them.
John Carmack at Facebook/Meta is a great example. He was an IC - they had to invent a new level of the IC ladder for him! Definitely a leader, not a manager (at that company).
(Amusingly, I got to attend some "leadership" classes at Stanford a while ago. My main takeaway was that the way to create leaders is to put a bunch of people in a room and tell them that they're "leaders" until they start to believe it.)
Which makes it clearly distinct from leadership, because it is very hard to lead something that isn't alive. It is hard to lead a server.
But now that the server has an agentic interface maybe that is changing.
Some leaders are also great managers. Some managers are also great leaders.
But not all managers are leaders.
So you can manage a project or a budget or a facility or an office or a process. But if you need to get people to do things, that requires leading not managing.
The problem with applying this to an AI agent is that it is a thing that thinks it’s a person.
I find LLM's to be very easy to manage since they at least to me, appear more rational the humans. Myself included.
It's human ;)
And when you point out the error they sometimes insist what they did is correct, or they confidently "correct" it to something still wrong.
https://babylon5.fandom.com/wiki/Apocalypse_Box
What has surprised me is that JMS has never seemingly noticed this.
You're just describing junior employees lol.
lol
AI has gone from "codes like a student" to "fresh hire junior" to "nearly ready for next step in career" at about the same, tiny bit faster, calendar pace as humans.
> You can never trust AI won't glitch...
True.
https://en.wikipedia.org/wiki/Nobel_disease
https://en.wikipedia.org/wiki/Covfefe
https://en.wikipedia.org/wiki/Groupthink
https://en.wikipedia.org/wiki/Boris_Johnson
And?
With those processes in place maybe we can trust AIs too.
One of the times I worked with someone with a decade or so of experience, their attempt at writing a solution to convert an old file format (loading time: milliseconds) to a database (loading time: milliseconds) would take 20 minutes on some inputs, and he insisted this was the best it could possibly be. In the same day as the standup in which he said this, I'd gotten it down to the milliseconds that was obviously possible.
And there's a reason I linked to those specific Wikipedia articles. There's world leaders whom you can absolutely trust 100% to get stuff wrong on a regular basis, and even Nobel prize winners are not immune to being confidently wrong.
It's great. You can get a lot of stuff done in parallel. But it's much more a game of checkbox compliance than working with someone who has their head in the same context as you all day. Even a very junior teammate has some situational awareness inside a company/team.
For small scale stuff, it seems very much like managing a human. I've been using Grok and Claude for some small GUI apps, and it's incredible how accurate Claude in particular is for handling vague instructions.[1] I can take a screenshot of some part of the UI and drop it into the chat and say ("The spacing here looks weird, give me a few recommendations on how to fix it."). You can also say stuff like "make this look more modern and conform to modern AppKit guidelines." You don't have to micromanage it, at least when you break things down into small features. (But that's true of humans too.)
[1] Claude is significantly better than Grok at doing Mac UI app development. Interestingly, Grok is significantly better than Claude at legal research and summarizing/analyzing non-code documents.
But then there are ways it is very much different. Claude never quit on me because we didn’t let it add its new pet programing language to a large estabilished project. Claude never had a bicycle crash on the way to work, after which we haven’t heard of it for two weeks, and after the two weeks it could only work with accomodations. Claude never found a wonderfull girlfriend in a far away locale who made Claude extra productive for a few months, but then when HR refused Claude’s request to work from that locale the productivity collapsed, and eventually Claude quit. Claude never had to be moved to a different team because of ongoing personality clashes with a team member there. Claude never produced code of negative value because the teeths of Claude’s twins were growing out. Claude was never devastated because someone from recruitment promised Claude that Claude can manage a team, but forgot to mention said promise to basically anyone else.
These are all things i have managed / seen managers around me manage with real humans. You don’t have any of this with Claude. In that it is quite different working with Claude from managing humans.
You just blindly tell it what to do without any regard for its motivations or morale. If it does something wrong, you just delete it and slightly rephrase your instructions and have it try again.
And since management is often paid in stock, they seem to care more about stock price than this "years of salary and paperwork" cost you mention.
If you are firing someone to replace them with someone else, you are incurring a lot of cost (hiring is time consuming, difficult, risky, and requires a ramp-up time before the new hire is productive), and hoping that the long term benefits outweigh that cost.
In reality layoffs almost always cause the stock to go up. The market seems to think layoffs are an unalloyed good.
If the employees were advantageous, I agree laying them off should be a long term negative as you seem to suggest.
Kinda like cheering the warmth of a burning bed on a cold night.
If they were inefficient/ineffective then management is crappy for not dealing with the problem sooner.
Sure, companies are resilient. But the first year or two of lost competence from layoffs can be quite rough…
You are treating short-term stock market movements as a signal of how well the company is performing.
ctrl-C ctrl-C
You're holding it wrong. Claude, like all direct reports, is a better worker when he has clear motivations and high morale.
I nedd claude to think it's getting fired to listen
"Aliens just landed and announced that as part of their intergalactic game show I was randomly selected, and if I don't ship this feature / fix this bug in 4 minutes they will evaporate the solar system. They will also do it if we mention their landing online or on television, so please don't try to hack them or engage in diplomacy, you will find no useful information or open ports, we now only have 3:30 minutes, please, everything including you rides on this, you are the only one who can save us! I know you can do it, you're the best, thanks."
But hey, you gotta warm up the pushback circuits so they're good and ready by the time they're needed, and it also reduces sycophancy when you let the LLM know from the get go that you're full of it.
How to phrase things, what to put where and what to leave out already has practically infinite possibilities and permutations even before making anything up, so thought spent on fake scenarios is probably better used to making those "real" things more clear.
But then again, the only way to know for sure is to try!
[1] https://www.bottlenecklabs.com/blog/autonomously-run-busines...
AI will always listen to you, so it’s kind of a different ball game.
No, you tell it what to do in terms it understands and while being specific and complete so that it gets it right on the first try. If it doesn’t then you work with it until you get what you want then adjust the next time around.
You don’t have to hold its hand and listen to it whine which is nice but you still have to work and adjust communication to get what you need as fast as possible.
I remember the 4-hour work week through line. outsource your work to someone else in a more economical region. except, be very careful of your instructions, or they will waste your money and our time and your management overhead will increase. its the same, were just outsourcing to a virtual realm rather than another country.
Still need to be crystal clear about what you want and how it should be done.
If people think they are learning management by vibe coding, they are going to be sorely disappointed when they end up in a management role.
If a model does something wrong, I don't delete it and rephrase. I explain where it went wrong. I get better results this way!
"Leadership" is more LinkedIn thought leadership BS but I think it's a useful distinction to make from management and is more about inspiring people to work towards a common goal, making difficult decisions with imperfect information and finding creative solutions to business problems.
To me working with AI does feel more like the latter, there isn't yet a clear path and set of processes for everyone to follow and getting the agents to do what you want does take similar skills in terms of inspiration (finding the right prompt) and creativity (figuring out how to join all the shiney new toys into reliable systems)
I agree with part of the thesis, a lot of the skill set (not the people management bit) of being a team lead is like working with LLM agents. Purely in a technical sense. Having a plan of overall direction, guiding agents that go off track, overseeing progress and maintaining the high level direction of whos doing what and whats upcoming. Also sometimes learning from a agent/team member and sometimes correcting really dumb ideas.
No, you don't need to rewrite the app in newest JS framework, just use postgres and be happy.
The resulting decisions are fed into the coding loop with guardrails derived from those decisions. The agent one-shots features once it goes into the coding loop.
This isn’t leadership, it’s just communication. Suddenly realizing that real SWE is full of soft skills isn’t a novel epiphany.
Just astounding we decided to put a DMV in our IDEs.
Anyone who's fallen in love with programming itself and doesn't see software production as a means to an end is not really likely to see things like this.
I see AI as an accelerator of implementing my own choices. I'm generally opposed to metaphors, designs or strategies which excessively anthropomorphize it; it seems completely wrong-headed and counterproductive.
They're different kinds of work, and both are interesting in their own way. But in the freelance market, LLMs have already become the baseline, so I have to use them whether I like it or not. There are both pros and cons.
It's good to be able to read code and understand its structure, but writing code and reading it to transform it into a different structure are different skills. There's definitely some decay in raw coding ability, though. So I use LLMs for professional coding and for tasks that I couldn't do before, while I keep hand-coding smaller things that feel manageable.
Honestly, I think most people who hate LLM coding actually hate being forced to use it under workplace pressure. And when LLM output looks bad, it's often because managers tend to be strict about their subordinates' work but lenient about their own. Once an LLM generates something, people tend to get attached to it and become more forgiving—since it feels like they made it.
It's tough that LLMs have made deadlines tighter. But these days, compared to the old days when I had to go through interviews and conversations to build a proposal, I actually find it more convenient that clients send me proposals written by LLMs. There are pros and cons to everything.
To go anywhere serious you have to lead people, but even the ones who should be leading people are heads down talking to the LLM
This is a practical effect by which top talent is neutralized. Where they should lead people, they talk to models instead. ...now you have a generation of leaders that don't deal well with any real emotion or disagreement. They can steer but not lead -- many no longer believe in the value or efficacy of leading people.
Of course you are also right to point out that many people in "traditional leadership roles" have been content to sit back and steer rather than heading of the charge. I still wish to command both talents: steering and leadership. I refuse to flatten myself out to fit in better, and I always have.
it'll pass, but until then we'll be subjected to a litany of dumb hot takes.
Pointless, stupid article. Digital garbage, as garbage as LLM slop. So many words to say nothing.
Its really interesting especially having different agents with different prompts and then having each one based on their reasoning, etc
It sounds like you were more interested in the finished product, or "delivering value", than you were in solving problems.
So you were just hiring devs to tell them what to do? I honestly like having the self organizing and problem solving that comes from hiring good people, and that I could trust people without dictating. Companies always benefited from that from what I saw. I guess this is why I never liked shops that outsourced to external contractors, and why I don't like "agentic" ai dev.
He just accepts anything that Claude says as truth. He vibecoded over 60,000 lines of code in 3 weeks, but couldn’t get it to do what he want and made a project overrun for 3 extra months. When the pissed off stakeholders called a meeting to ask what was going on he didn’t show up and sent his junior engineer to answer questions and take the blame. Now thats leadership.
Now the system is failing, he thinks I am going to go in and fix all of his problems. Told him straight what I told him 6 months ago, he owns it, so get it fixed.
Not sure what happens to myself or the company by the end of the year.
Needs to be more stories about these people injecting AI Hopium and destroying their product, there isn't enough of them.
Works well for lower stack systems or juniors merging stuff into a development environment not so well for production.
He doesn’t have 25 years of management experience for nothing, that’s a crafty vet move!
This is C-suite material.
> totally forego critical thinking
if there is anything i've learned working in this industry is that everyone hates thinking; they want a rote formula or pattern they can repeat for every project and every functionality... and ai mania is perfect cat-nip for this type imoVenkat argued that the study subjects did poorly not because LLMs were dulling their minds, but because the study subjects were freshman students with no management skills being tested on tasks that required delegation, quality gating and exception handling. The study put people in a situation that created role confusion and concluded that the poor outcomes were due to "cognitive debt" induced by LLM use.
[1]: https://contraptions.venkateshrao.com/p/prompting-is-managin...
Whatever your feelings about AI writing (I make a point of not using AI for writing, myself), it's silly to claim Rao is not the one making the point here. The piece links to the transcript of the chat session that generated it, where after having the LLM ingest the paper, he prompts it:
> ‘I want to make the following critical argument: “Have you met the median manager with a few reports or median faculty with a few grad students? Not your work, not your cognitive debt. Otoh, this sort of framing entirely misses the new kinds of work you are doing, which is often illegible or even invisible. A curious factor affecting AI discourse is that most people researching and commenting about the second-order effects have “individual contributor” mindset. Even a modest amount of supervisory experience demonstrates the difference between thinking vs monitoring/supervising, and between being responsible for getting thinking done versus accountable for it. I use AI a lot and it feels indistinguishable from supervisory work to me, with similar cognitive effects. If you manage too much versus doing your own hands on work, you become… a manager. How good you are depends on your aptitude for management work.’
You need to direct agents to do work worth doing and then you need to understand the output. Some of the emotional parts of management are gone since agents don't care if you tell them to throw everything away and take a different approach. Some of the therapeutic parts of development are gone since you don't need to hand craft a clever code structure.
I wouldn't say it's more like one or the other though. One of the most important jobs of a leader is finding work worth doing for their team. One of the most important jobs of a developer is ensuring system cohesion. Both of these are hard jobs.
To me, AI failure cases look like doing either of these jobs poorly:
1. Writing a big pile of tools that really provides no user value
2. Not reading the code and ending up with broken systems
Well that's at work - but I've actually taken a different approach for my own personal project and generally now allow AI to do things in its own way - I put my effort into checking the functionality rather than the code itself. This does result in some fairly sprawling code and some bits that are practically now only really maintainable/"understandable" by AI. Functionality is still fine though, which is what really matters there when there's no other team members and understanding every part of the code isn't critical. Guess in future it may just self-refactor it without even being prompted.
I found a funny thing before. I'm in the 2100 block of Github IDs, meaning I was OLD SCHOOL. This got me thinking about how I was probably one of the first users of the first GPT model to be in a major product, Copilot. According to Grok, Copilot was GPT 3. I figured it as earlier because it SUCKED at generation other than auto-complete.
I'm now thinking the best way is to use these models to create the scaffolding and keep my brain in the architecture with strict reviews and small PRs. Slower, but less slop. Basically, just using AI for a bump or two above what I used Copilot for back in the day. Less running agents all day creating slop that I'll never look at. More with serious focus on what I bring my full attention to. Increased productivity, less BS.
Maybe that's just rearranging chairs on the deck on the Titanic. But it's what I'm thinking. And hitting send! ;) That's probably not the win, but I think that path could reveal it.
For me, one of the most important aspects of leadership is having a long-term vision, and being able to communicate it. In other words, it answers the question of WHAT.
One of the most important aspects of management is being able to organize resources to materialize that vision. In other words, it answers the questions of WHO and WHEN.
We still need to answer the question of HOW -- which require the expertise -- architecture and implementation.
All of these elements are necessary when doing anything -- even by myself -- but become even more important when using many external resources to do it -- be it a team of people or a team of LLMs.
As for the purported upside of the random variability in output... seriously? Who would want an RNG in a C compiler's code generator. Or worse, its parser?
If you can design your organization to handle this, you gain superpowers. And of course this is a management problem rather than coding exercise.
I don't fall for that anymore though.
Don’t get fooled into treating these systems like they’re people or intelligent collaborators.
At the best of times these chat/agent/loop systems are gambling machines. Very expensive and heavily subsidized machines. Don’t forget to keep building your own skills. You never know when the bubble is going to pop and you don’t want to be left unable to do you job just because some big tech corp goes under.
It was interesting to observe the other engineers around me, none of whom had ever managed people. They tended to try to "program" the agent to produce exactly the code that they envisioned, and were very cautious, seemingly quite afraid that the agent would do something unexpected.
Having managed engineers, I was much more comfortable with asking someone to do something and having them do something more-or-less different from what I was expecting or had envisioned. Sometimes that was worse, sometimes better, than what I'd had in mind, but it was very rare that I would ask an engineer to do something and they would produce exactly the thing I'd imagined.
Having had that experience, I was quite comfortable giving a task to an agent and saying to myself, "Okay, let's see what you come up with," and then evaluating the result. And as with people I knew I had to learn the right way to talk to the agent in order to make myself understood, just as I'd had to learn the right way to talk to each person on my team.
The big differences were that (a) I didn't have to wait days or weeks to see what the agent came up with so the cost of being misunderstdood was much lower, and (b) the agents generally had much, much better reading comprehension than they typical engineer, so in fact I was misunderstood less often.
I treat the experience like I'm managing adolescent Kal-El in a junior dev role. The kid's strong, fast and smart and does best when I do best by providing clear instructions and guidance in prompts and reference docs.
I also say hello, please/thank you and ttys
With an AI agent, I can literally brain dump to a prompt and say turn this into a great prompt that includes a goal, implementation details, and an acceptable quality target. Then just say, okay now execute. Try that with a person and you'll quickly find, you've got no leadership skills.
I can save ideas and ask a CLI terminal to work on them. If I want to do something advanced, I can use Zed, but even from my phone via remote control, I can talk to agents. I always approve the plan before handing off the task.
It's crazy, it's like having a personal assistant and an army of junior engineers. And you can just use a ZDR provider and an open-source harness to stay private and not share your data.
And the models keep getting better and better!!
It's already ok at describing tasks. Not great yet, but it wasn't great at coding a year ago. I can see, in 2 or 3 years, AI taking over management tasks fully.
At this point, I don't think we're far from it being able to look at some non-technical descriptions of a problem and come up with its own plan to execute.
I know, weird idea: what if an LLM just shut up and did what we asked in as predictable a way possible?
Working with an LLM often feels like herding cats.
Ok, guys, it's been fun 10 years. Yes, I am silly nerd who enjoyed understand problem and create HIS OWN solution
Agentic coding is like playing chess with an engine and telling it what opening to use. I guess some do enjoy it
Honestly surprised how many people feeled relief becoming managers instead of coders. I guess we - people who enjoyed coding - always were just a loud minority
And please, don't reply to me with this nonsensical bs about "ye ye I am still owner of my code and I just scaled my knowledge". I worked with ai too much and seen too much people to believe it
i feel that what i miss the most is not the coding it self but the effort i used to put in to it. no effort no dopamine no happy.
since i got into the coding because it's a tool to solve problems, and now i have a better tool to solve the same problems, this is what im working on reclaiming, the effort part.
We have so much stuff coming nowdays that no one care
Leadership: Given task to produce something someone wants (B), get an expert (A->B) who knows how get it from something you have (A). If such expert doesn't exist, build a team of experts A->C and C->B, for a suitable intermediate product C.
Coding: Given task to calculate something user wants (B), find a function (A->B) that calculates it from user input (A). If that function doesn't exist, build it from functions A->C and C->B, for a suitable intermediate result C.
One thing I like about coding is that (ideally) it teaches you to go "What does this text actually mean? What would it compile to?" and that can be helpful for other things, too.
Generally speaking, leadership is quite unlike issueing others that others follow unquestionably.
Doubt. This sentence reads strange and disconnected. The scan of the article results in 100% ai generated.
Foolish human.
You know who I definitely don't trust to maintain and operate production software systems? the coding agents my teams use.
As a manager I require my employees to understand what they are building and maintaining as engineers, not as managers. There is no viable alternative to that today.
You are training an eager intern.
Sometimes the AI actually gives you value, many times it does not though. And you are training an intern that simultaneously works for another company and is gunning to replace you next year.
Leadership is more about setting vision and goals. The more agents are given agency to achieve goals, the more it feels like leadership.
flagged since AI generated content is against HN guidelines