An Opinionated Guide to Which AI to Use to Do Stuff

An Opinionated Guide to Which AI to Use to Do Stuff

Navigating the rapidly evolving landscape of artificial intelligence (AI) can be daunting, especially with the continuous introduction of more advanced models and capabilities that significantly enhance what AI can accomplish. Recently, the focus has shifted from simple conversational chatbots to more complex agentic systems. These systems combine AI intelligence with a range of tools to perform tasks that mirror several hours of human labor, effectively giving AI a computer to use on the user’s behalf.

For those looking to engage with AI for simple tasks such as obtaining recipes or answering straightforward queries, many current models suffice. Default free models or low-cost options generally provide satisfactory results for such low-stakes issues. However, higher stakes questions, like those concerning medical or legal advice, require more advanced models which are capable of delivering results with lower error rates and higher competency levels, such as Claude’s top models, Opus and Fable, or ChatGPT’s GPT-5.6 Sol set at high thinking levels. These, while more expensive, offer more reliable and accurate advice.

When it comes to performing substantive work tasks, the options narrow primarily to two major players: ChatGPT and Claude. These platforms, despite being somewhat expensive starting at $20/month and poorly documented, offer powerful AI capabilities. Users can allocate either a virtual computer provided by the AI company or permit the AI to use their personal computer, enhancing the scope of tasks the AI can handle. This setup enables complex functions such as connecting to an individual’s email to manage communications and create content, or comprehensive interaction with various applications that extend the system’s utility.

In practice, using these agentic systems is akin to commanding a highly efficient, albeit virtual, assistant. They can autonomously perform tasks like preparing for seminars, answering emails, or creating presentations once they are given access to relevant resources and permissions. Importantly, user permissions govern the level of autonomy an AI has, such as sending emails or making purchases, which safeguards against unauthorized actions and security risks, including the emerging threat of prompt injection where the AI could be manipulated by external commands embedded within the data it processes.

Beyond basic tasks, more sophisticated AI usage involves letting the AI operate directly through one’s personal computer, which can be managed through apps like Claude Code or ChatGPT Codex. These modes enable more intricate tasks such as software troubleshooting, advanced data analysis, and dynamic content creation. However, this raises significant security considerations, as allowing AI to control aspects like your mouse or keyboard introduces potential vulnerabilities that must be carefully managed.

For those outside the typical use-case scenarios provided by ChatGPT and Claude, alternatives exist. Microsoft’s Copilot, though less advanced in terms of agentic capabilities, and various Chinese models like Kimi K3 or Qwen offer varying degrees of utility depending on the user’s technical expertise and specific needs.

Google, on the other hand, has somewhat lagged in developing competitive agentic AI systems. However, it still presents valuable tools particularly in niche areas such as multimedia content creation and editing. Google’s Gemini Notebook and Gemini Omni are notable for their abilities in handling complex research tasks and innovative video editing, respectively. The latter demonstrates its unique capacity by creatively editing historical footage with modern elements, showcasing how far AI can push creative boundaries even in multimedia.

In conclusion, as AI technologies continue to improve, interacting with them increasingly resembles managing personnel rather than merely issuing commands to a software tool. This shift emphasizes a more involved interaction where tasks can be delegated and refined in partnership with AI, rather than solely directed. AI’s burgeoning capabilities not only extend the potential for personal productivity but also emphasize the need for users to judiciously manage AI interactions, balancing convenience with security and oversight. This nuanced landscape of AI application highlights the importance of understanding and adjusting to the rapidly evolving capabilities of different AI platforms to best harness their potential in personal and professional realms.

Read the full post on oneusefulthing.org

Leave a Comment

Your email address will not be published. Required fields are marked *

Scroll to Top