Where could “AI” be actually helpful?

I don’t know… for me, learning the camera’s settings is half the fun. I mean, I like taking pictures as much as anyone, but the learning a new camera is a fun exercise in its own right. If I wanted to just point and shoot, I’d just set it to Auto and go about my business.

And I’m not talking out of my ass… I recently bought a Canon EOS R, which admittedly isn’t as modern as a Nikon Z8, but it’s a modern mirrorless camera nonetheless. And since I came from an old Canon T2i, it’s a lot of fun seeing how it works differently.

FYI… one cool thing with modern mirrorless cameras is that since they’ve got much shorter focal lengths (i.e. the distance between the back of the lens and the sensor), you can get a simple adapter that adds a few millimeters between the back of the old lens and the sensor to line it all up. So all those old-school lenses from the 60s through the 80s are now an option. I recently got a Konica Hexanon 50mm f1.7 for $10, and it’s a lot of fun. Especially considering the utility of modern mirrorless focus peaking type abilities.

If you use a single claude code instance for coding, planning, debugging, and analyzing results, it’s not as powerful. It’s better to split the cognitive load among parallel claude code instances for the same project - one is a coder (ONLY writes the code), one is a planner (writes specs, sprint maps, etc. for the coder), one is a debugger (looks at log outputs, failure modes, etc.) that writes reports for the planner or coder, one is an output analyzer (looks for correctness, does this match the output the planner expects, etc.), The specifics of what parallel instances you might need depend on the project. If you avoid cluttering the individual agent context windows with mixed domains (writing code vs debugging vs planning what to do next in the code or how to fix a bug you found), individually each instance is much more powerful. In this model, you are as much an orchestrator as anything else, just applying human judgement as needed across the set of agents.

I normally do start a different session or sometimes even switch to another model entirely when debugging. I still often run into instances where the LLM can’t seem to figure out what the problem is, and it instead presents a very detailed explanation that sounds very plausible but unfortunately does not fix the problem.

It’s fun when you have an actual printed manual. I used to keep manuals in the bathroom for reading while I was doing my business, and I would read that thing cover to cover. When it’s only available online, it’s not as convenient. Or I can download it and print all 74 pages myself. Then there’s the Z8 reference guide, which is a 1000+ page PDF or available online.

Actually I think that considering their amazing language proficiency, including the ability to read, understand, and summarize vast quantities of written material, ability to comprehend images, and ability to generate images and video and synthesize speech, LLMs are the most general-purpose AI model yet developed.

Very true. I’ve used ChatGPT and Claude to help me remember a book, film, or short story based on only a very tenuous description. In one case, I had been Googling in vain for quite some time and finally decided to ask ChatGPT. It identified what I was trying to remember in a mere second!

But this is just a special case of LLMs generally being a very powerful interactive information resource with which you can carry on a productive conversation and refine your knowledge, and explore concepts and ideas. One has to be cautious of the fact that it can get things wrong sometimes, but the capability is so powerful and groundbreaking and since I find that it’s right most of the time I find myself using it a great deal for just that purpose.

Which sounds a lot like a human debugging session to me.

Well, the thing about that is: In the end, I get frustrated and look at the code myself, and I can normally figure out the problem and fix it in about 30 minutes.

I probably just need to realize after a second failure of the LLM at a debugging task I should cut my losses and look at the code myself. But in my experience a LLM is far from a panacea when coding.

I use the YouTube Gemini summary function to distill a 30 minute video into the short summary of the highlights, which is all I mostly need anyway. Then I can skip to the next video and not have to wait for the bullshit like number 10 reason will surprise you ruse.

Is the summary in the form of writing, or video outtakes of the 2 minutes of actual content in the typical YT vid?

I think replacing child actors with AI actors is at least worth considering.

FYI, this is worth a look. Where it differs from some silly parlour trick is that the transformation of the singer’s face into Simon Cowell’s face is absolutely perfect, and note that this is being done in real time as the singer performs.

Yeah this happens too, and sometimes the LLM makes it more difficult by patching over a real bug with a kludge overfitted to the specific error that papers over the root issue, making the root problem even harder to find until you look manually. Learning the nuances of how the coding tools work, make mistakes, hallucinate, etc. helps a lot - the problem here is that the tools are changing almost as fast as the learning process takes to use them to maximum effect.

Its two or three short paragraphs. It generally give a synopsis of the video, lists out the key points and then summarizes. It really beats watching a video that drags on for 20 minutes that should have been a 5 minute one.

I then can choose to watch the video or just go to the key point I’m interested in. Or not watch it at all. :slightly_smiling_face:

I think this really depends on the context. Watching AI children get victimized is just as bad as real ones. It could give sick people ideas and normalize it to them.

What is the issue that AI actors would solve that tighter legal controls and actual punishment could accomplish? Overwork? Abuse off film by adults in power? These things happen to adults as well.

Oh man, the weird kludges are the most frustrating about coding through AI. The aforementioned silent 60 second timeout on a function was so irritating when I found it, I had to go outside and take a break. Nobody asked for it, it never mentioned it in its plan, and it certainly wasn’t a choice I would have made. I stopped auto-accepting its edits and reviewed the changes before I let it proceed after that.

That’s true, and they can be weirdly uneven on specific tasks. When feeding the code by one model into a different model when trying to debug it, the new one will sometimes offer up a new set of code for a completely unrelated function in addition to its proposed fix. I’m not sure what triggers that odd behavior. I seem to agree with the new LLM and accept the rewrite of the function about %50 of the time. A few times the new suggestion has been batshit crazy, though.

I won’t argue. And given that they are getting rather good at things like image recognition and creation, it’s not clear that LLM is even the best description? Under the covers it’s all pattern recognition and extrapolation, I suppose. Language or otherwise.

But is this the (or a) path to something we would call strong general AI? I really don’t know, how would we tell?

Unless you are religious, or a quantum sceptic like Penrose, it seems that mind arises from some sort of emergent process?

It turns out that a truly deep understanding of language – which LLMs clearly possess – necessarily begins to model basic human cognition across a surprisingly wide spectrum of capability, so that things like image processing and generation are not as much of a stretch as they may seem.

No. LLMs are just one particular breakthrough technology. True AGI will need to be a combination of many others, with LLM language proficiency being basically a very good front-end, or UI.

It’s all about emergence at sufficient scale. Because there’s nothing else there, otherwise one has to believe in magic, spirits, immortal souls, and fairies.

All right, I’m going to share my experience here, though this may be an unusual application of AI. This is long, sorry. But it goes unexpected places.

I started experimenting last week with having GPT give me strategic feedback on written drafts of grant applications. I put a lot of time into it - this is not time-saving in any respect. For context, I have fifteen years of experience writing grants and consider myself a serious writer. While GPT is often more concise than me, it is not a better writer. But what it can do is point out places in the Notice of Funding Opportunity that don’t align with my proposed project, and make suggestions for integrating funding priorities from the NOFO in specific places in my narrative. In a pinch, such as yesterday, it helped me cut 250 words from a narrative.

So based on that limited success, I asked it to create a self-guided course for me to learn how to integrate AI into my grants management workflow. By this point, it knows everything I’ve told it about the nature of my job, how long I’ve been working there. I asked it to ensure the coursework was ADHD friendly, and I provided it with an overview of my cognitive strengths and weaknesses, and this is where it took a turn. Suddenly the project focus veered from improving AI skills to addressing my specific executive functioning challenges within the context of my specific job.

Now I’ve read books broadly about how to deal with ADHD, and I’ve read books broadly about how to be a grants manager, and I even had an actual ADHD coaching group, but the level of specificity I got into with GPT was unrivaled. After some back and forth, GPT said, “I notice you said you miss things frequently. But elsewhere you say you’ve never missed a deadline. It seems like the reason you’ve been so successful is because of constant hypervigilance and all the tasks you have to keep in your head - You’re not forgetting things, you’re remembering everything all the time. This is probably why you feel so overwhelmed. So let’s create a strategy to reduce the burden on your working memory so you only have to hold one thing in your head at a time.”

I’m not gonna lie, I almost burst into tears. I’ve been well aware of the problem for years, and have tried countless things to solve it, including using Trello. So I told it about Trello and it advised me to stick with Trello, but with some tweaks. We created a category called the Parking Lot where I get to put everything likely to distract me, including creating new workflow systems. We talked about what, on Trello, I found overwhelming. And we were able to collaboratively hone in on a key insight: when I see the name of the project, I am immediately holding all of the project in my head at one time. GPT suggested I try highlighting next actions instead. It suggested I put a checklist in my Grant Application Trello card (duh) but this led me to the insight of putting the next action right there in the card title, which led to a whole different insight:

Specific action step first. Project second.

So instead of a Trello card reading: “Start OVC Services Grant,” the Title of the card now reads something like:

“Fill out grant revision matrix: OVC Services Grant”

This makes it so my brain registers the specific, small action step rather than the whole project at once.

I also worked on tweaking all my Trello cards in this format and operationalizing the task titles that felt vague or overly complex. So a task called “Look into HUD CoC grant” that I’ve been avoiding for literally months became “Find most recent NOFO: HUD CoC grant”

How hard can it be to find a NOFO?

Every time I complete an action step, I cross it off the card’s checklist and edit the title with the next action step. And so in this way I am only ever looking at the next step.

It’s bloody genius.

We worked on some other tweaks over the course of about four hours. What’s most impressive to me about this whole thing is that after I made these changes, I got three or four tasks done right away. I submitted a grant report, we created a revision matrix using the specific rejection feedback I got, we created a matrix template for future grants, and we created an Excel workbook to document general learnings from grant reviewer feedback.

It seems like the potential for integrating AI into the grants management workflow is pretty much unlimited, but right now I’m focused on developing what GPT calls my “working-memory-reduction-system.” Since I told it I have a tendency to hyperfocus on lower priority things, it told me to put working on my system in the Parking Lot until I complete my highest priority tasks.

I have deep existential thoughts, doubts and philosophical questions about this (and I think brain fry is a real thing after spending hours in constant dialogue with an LLM) but the bottom line is I think this project is going to help me a lot, in ways I didn’t expect.

It is also the standard guidance from several effectiveness practices. Don’t make a line item in your to-do list that is bigger than (e.g.) one hour’s work or that has any prerequisite you haven’t already fully completed.

This is one of the biggies along those lines.

Can recommend.

I’ve read that, and I’ve used some aspects of GTD, such as having an inbox where everything is collected. Maybe I just sucked in the application of these principles. My Trello board was an improvement but I still felt kinda overwhelmed/confused. It was useful to get specific feedback on how my Trello board was structured. I’ve probably read at least 200 books on productivity in my lifetime. I have tried I cannot tell you how many productivity apps. And the how-to-write tasks thing never clicked.

Oh, I’ve heard to break it down into smaller parts, but what I would then have is a long list of smaller parts which was, in and of itself, overwhelming. It never occurred to me to have only the next task visible at any given time.

It’s not so much the access to information I lacked so much as the understanding of how to apply it to my specific workflow. And I had never heard it framed in terms of working memory and reducing the burden on my working memory. But when I look at what I was doing, yeah, no wonder I felt so overwhelmed all the time.