What about giving it to your secretary with a list of what to shop?

> What about giving it to your secretary with a list of what to shop?

"When a Comet user directs the Assistant to locate an item on Amazon.com, the Assistant takes screenshotsof the browser view, sends those screenshotsfrom the user’s computer to Perplexity’s servers, and receives instructions from Perplexity’s servers on how to navigateAmazon.com. In other words, the Assistant cannot operate wholly independently; it relies on direction from the user and instructions from Perplexity’s servers."

Obviously all browsers rely on direction from the user. But they also typically rely on instructions from the browser maker's servers. Traditionally, you get all of those instructions in a single download that's been pre-packaged (the browser program itself). But if part of the browser's logic requires more computational resources than most consumers have, what's the problem with "outsourcing" that bit to Perplexity's servers?

As a consumer, I don't think it's very prudent to trust a company with that kind of access to your data, but it doesn't seem materially different from the kind of trust you have to give to Chrome.

> if part of the browser's logic requires more computational resources than most consumers have, what's the problem with "outsourcing" that bit to Perplexity's servers?

It's not longer my agent. It's a joint agent of my self and whoever else is giving it instructions. (This is Amazon's argument.)

Also, Amazon told Perplexity to fuck off. If your assistant is told to stop doing something in a cease and desist, and then keeps doing it, the fact that you asked them to do it isn't a get-out-of-jail-free card.

My guess is we'll find a balance where local models can act as user agents while corporate ones have to meet certain requirements to gain that safe harbour.

[dead]