How much do you trust AI agents?

With the advent of clawdbots, it's as if we've all lost our inhibitions and "put our lives completely in their hands."

I'm all for delegating work, but not giving them too much personal/sensitive stuff to handle.

I certainly wouldn't trust something to the extent of providing:

  • access to personal finances and operations (maybe just setting aside an amount I'm willing to lose)

  • sensitive health and biometric information (can be easily misused)

  • confidential communication with key people (secret is secret)

Are there any tasks you wouldn't give AI agents or data you wouldn't allow them to access? What would that be?

Re. finances – Yesterday I read this : Sapiom raises $15M to help AI agents buy their own tech tools – so this may be a new era when funds will go rather to Agents than to founders.

4.9K views

Add a comment

Replies

Best

Great question. Trust in AI agents comes down to one thing: can you see what it's doing and stop it if needed?

We're building AnveVoice — a voice AI agent that takes real actions on websites (clicks buttons, fills forms, navigates pages). The trust challenge is huge because it's not just generating text — it's actually interacting with the DOM.

Our approach: every action is transparent, reversible where possible, and the user stays in control. Sub-700ms latency so there's no lag between command and action. WCAG 2.1 AA compliant so it's accessible to everyone.

The key insight: trust scales when the AI operates within clear boundaries. We use 46 MCP tools via JSON-RPC 2.0 — each tool has a defined scope. The agent can't go rogue because its capabilities are explicitly defined.

MIT-0 licensed, free tier available at if anyone wants to try it.

 Thank you for announcing this option.

It depends on the context, but they should always be double checked.


When it comes to scraping & data retrieval tasks, they usually perform fairly well.

On the other hand, for creating proper marketing materials, you have to do extensive checking & adjusting to get the result you want.

This can also play a factor in trying to determine if AI agent has produced something that's actually real or fake -> this is because once they're tweaked significantly, they can be extremely deceptive.

This is something I'm tackling in the shopping space with my launch today - we are currently ranked #4 on PH!

 I liked the launch tbh :)

I wouldn’t delegate anything related to payments.

I’d only allow access to things that are already paid for.

Leaving payments open feels like a really bad idea to me.
I’d mainly use AI agents for tasks like market research or summarizing things that require regular updates.

 payments are no no to me as well ;)

About the same level of trust I have when I leave my 3 teenage sons at home for the weekend.

 :D Teenage vs AI agents. I would trust AI agents more in this case :D I know teenagers :D

I trust AI agents for routine tasks and productivity, but not for sensitive decisions or personal data, human judgment and verification are still essential for reliability and safety.

 I perceive it in the same way :)

I kind of come at this a little differently. The term "AI agent" can mean different things to different people. Yes, there are lots of people going all in on OpenClaw. It's definitely not ready for average human use. You have to really be savvy to use it safely. Most people won't take care, and will yolo and maybe regret.

What I have changed fundamentally in my every day life is I use AI tools interactively all day and sometimes turn them into repeatable tasks when I've honed the skill and trust it. I give Claude access to much of my work and personal data through well governed and scoped MCP's I control where I can log activity. The main thing I'm trusting is that Anthropic isn't slurping my data and I make sure I've configured Claude to restrict from web search and don't enable their stock MCP's.

Once you start prompting your way to simplify your daily work or life tasks, it's a drug of productivity. Literally hours of work reduced to nothing. Once you trust the results and skill them up as scheduled activities, you can't imagine going back. I recommend everyone I work with to start with AI through chat tools; figure out how to prompt better, establish guardrails, save off skills. It's a necessary skill in 2026 and preps you for what's coming down the pipe as the tools get better.

 It would be useful in that case to find a good way or process to track and control the work of those agents effectively, so that there is no significant irreversible harm. :)

 for sure. trust builds from usage and expands with the visibility of tracking what the AI has done. indeed.

Honestly, I’m a bit concerned about the idea of a blackbox. I’d prefer it to be more transparent, of course, as long as the content doesn’t end up feeling like spam.

 Transparency is good but not enough tbh, I require more controlled environment.

It really depends on the domain for me. I work in HR tech and we use AI agents to filter noise and surface patterns that humans would totally miss. But the final call on a candidate? That still needs a person who understands context, culture fit, and empathy. Trust builds when the agent is transparent about what it did and why, not when it just hands you an answer. The biggest risk isn't the AI being wrong, it's people not questioning the output.

 this is interesting – how do you use AI for your work and how do you see candidates using it? Are they relying too much on AI?

 I use AI every day. I've trained my Claude on my tone of voice and it knows about my company and what we stand for, so a lot of admin work is eliminated.

As for candidates, I don't think using AI is a bad thing at all. I actually see it as a positive signal when someone knows how to leverage it well. But they should be mindful about making the output their own. If your resume or cover letter is full of em-dashes and words like "spearheaded" or "busywork," it's immediately obvious it came straight from ChatGPT. The suggestion would be to take the time to adjust it so it sounds like you, not like a prompt.

Photos and finances are my hard no. Happy to delegate tasks, but not trust AI with anything deeply personal or financially sensitive.

 Mentioning photos is new in this discussion. Why so?

I run 10+ autonomous agents daily for PM workflows. The trust framework I've landed on: read access is liberal, write access is reviewable, anything touching external systems or money is approval-gated. The paranoia goes away once the boundary is clear.

 How much do you pay for running it? :D

 they are all using Claude 200$ subscriptions. And this is the beauty. One subscription to chat in Claude, use Claude Code and run agents based on Sonnet and Opus :)

First
Previous
•••
111213
•••
Next