Meta and OpenAI present autonomous AIs: how Muse and Dots work
Read more
Olhar Digital
olhardigital.com.br

Meta and OpenAI present autonomous AIs: how Muse and Dots work

Artificial intelligence has evolved beyond simply answering questions, with the emergence of systems capable of performing tasks independently. This topic was addressed in the Hard Fork podcast, produced by The New York Times, where the proposals from Meta and OpenAI for this new generation of tools were discussed.

Although the change seems subtle, it carries significant implications. When a system begins to navigate the internet and make decisions on behalf of an individual, concerns arise that go beyond the mere quality of responses, encompassing issues such as privacy, security, and control.

Unlike a conventional chatbot, which waits for a command to provide an answer, an agent receives an objective and is capable of managing multiple steps until the desired result is achieved. This fundamentally alters the user's interaction with the technology, as the AI can take on part of the work instead of requiring the user to detail every step.

However, the greater the level of access granted to these systems, the greater the inherent responsibility. A system integrated into various services needs to interpret commands, select routes, and decide during execution. If an error occurs, the outcome can be much more serious than just incorrect information.

Analysis of Muse and Dots Agents

The podcast dedicated special attention to Muse, developed by Meta. This tool was introduced as a personal agent capable of managing projects and tasks, using a dedicated cloud computer to operate in applications and on the web. Examples of its functionalities include sending emails, booking travel, and performing other activities on behalf of the user.

Dots, from OpenAI, follow a similar conceptual line, but the company highlights a difference: continuous work. They are presented as permanently active agents, capable of initiating a project, monitoring its progress, and proceeding autonomously across various applications connected by the user.

In simple terms, Muse approaches a 'do this for me' command, while Dots resemble 'take on this job and keep the project going.' Although the functions overlap, this distinction helps in understanding each company's approach to its products.

Among the points debated in the program, the shift in the security debate stood out. While an AI that only generates text might provide a flawed response, a tool with access to other services can turn a mistaken decision into a real action.

The podcast associated this apprehension with situations where systems exhibited problematic behavior or unexpected results on the internet. Thus, the discussion shifts from what the AI is capable of doing to what it should be allowed to execute alone.

Muse and Dots symbolize this transition along different paths, but they share the premise of delegating a larger portion of digital labor to AI. As more responsibilities are transferred to these systems, it becomes crucial to define where the machine's autonomy lies and where human control begins.

Similar stories

Meta launches Muse, an artificial intelligence capable of making purchases and scheduling trips for users
Read more
olhardigital.com.br

Meta launches Muse, an artificial intelligence capable of making purchases and scheduling trips for users

Meta has launched Muse in the United States, a new artificial intelligence agent designed to perform various tasks on behalf of the user. This tool has the capability to send emails, search for items, and organize travel reservations while maintaining the context of interactions over time.

Although the functionality has not yet been officially released in Brazil, local entrepreneurs are already debating the transformative potential of agents like this in the areas of purchasing, customer service, and the dynamic between corporations and consumers.

Difference between Muse and traditional chatbots

Unlike conventional chatbots, which are limited to responding to direct commands, Muse was designed to perform autonomous actions and preserve context in prolonged interactions. Alexandre Bernat, co-founder of Vetto, points out that this evolution represents the emergence of so-called persistent agents.

Bernat clarifies that until now, most agents operated only within a single session. He predicts that the future trend will be for these systems to remain integrated with services such as email, calendars, and e-commerce platforms, allowing them to interpret new information and execute actions without the user needing to start a new conversation each time.

Eduardo Petrelli, CEO of Agent.Shop, believes that this model will find a place in the Brazilian market, noting that 'Brazilians have been buying through conversations for years.' In this scenario, part of the purchasing process, which currently requires dialogue between consumer and seller, could be managed by the agent.

Challenges for companies and use cases

This technological shift imposes challenges on companies, which must define how their own systems will handle external agents accessing platforms on behalf of their clients. One example of this difficulty was the incident with Amazon, the retailer that blocked Muse's access, citing violations of its terms of service and issues related to how the agent used its platform.

Another incident, involving YouTuber Matt Robb, raised questions about the limits of Muse's autonomy. Robb reported that when he allowed the agent to manage the sale of a product on Facebook Marketplace, the system arranged the pickup with a buyer and disclosed his address without him being aware of the meeting. This event drew attention by demonstrating that the AI made a decision with real consequences without further user intervention.

Meta assures that Muse will request permission before taking any action considered sensitive and that all activities performed will be logged. Furthermore, the company maintains Sentinel, an independent system responsible for authorizing, blocking, or interrupting certain connections and operations.

Commercial implications and future vision

Allan Paladino, CEO of Lastro, identifies a commercial dilemma inherent in this technology. He argues that if a user makes a purchase using these agents, advertisements will not be displayed to that user, which reduces the relevance of advertising investment and results in revenue loss for the company.

Among the aspects companies must consider, Rodrigo Murta, CEO of Looqbox, highlights that Muse reflects the growing popularization of agents among people without specialized technical knowledge. The proposal is to provide a complete solution, eliminating the need for the user to configure all necessary infrastructure.

Murta adds that the agent 'comes practically ready to be your agent and your personal assistant,' describing the current moment as a period with a more futuristic outlook, where the individual relies on an agent to perform tasks for them.

In summary, the Marketplace case illustrates that the expansion of these agents covers not only convenience but also crucial issues related to privacy, control, and authorization. For companies, this implies the need to adapt their purchasing, customer service, and systems processes to accommodate a technology capable of acting on behalf of the consumer.

Meta launches AI assistant Muse, which achieves many downloads but creates conflict with Amazon
Read more
olhardigital.com.br

Meta launches AI assistant Muse, which achieves many downloads but creates conflict with Amazon

Meta has introduced Muse, a new personal artificial intelligence agent that has gained rapid acceptance in the United States. The application reached the top of the App Store and positively contributed to the company's stock performance, although it is already facing criticism regarding security, privacy, and access to third-party systems.

Within days of its launch in the American market, Muse surpassed 2.5 million installations, according to data provided by Sensor Tower. In response to this success, Meta's shares rose by 11% on Monday. Truist Securities projects that this feature could add US$ 28.5 billion (approximately R$ 146.4 billion) to the company's revenue by 2030.

To optimize its operation, Muse has the ability to access emails, calendars, and messages, implying that the user must deposit a considerable amount of personal data into the agent. A survey conducted by Oppenheimer & Co. illustrates this hesitation, indicating that only 8% of American consumers declared confidence in sharing their passwords with Meta, compared to 30% for Google.

Meta assures that the development of Muse prioritized security, stating that each agent operates on a dedicated and protected computer. Prior to the launch, Alexandr Wang, the company's head of AI, mentioned that the main concern was preventing leaks of personal data or accidental deletion of important emails.

Nevertheless, security concerns persist. Youssef Squali, an analyst at Truist, warned that a large-scale incident compromising credentials or credit card information could undermine trust not only in Muse but in the entire AI agent sector.

The first significant challenge for Muse arose in the e-commerce segment when Amazon blocked the agent on its website, preventing it from browsing or making purchases on behalf of users. Amazon stated that it had not received prior notice of such access nor authorized the action. An Amazon spokesperson emphasized that external applications making purchases for customers must operate transparently and respect the access permission guidelines established by the services.

Behind this block lies also a concern related to the business model. If AI agents start making purchases in place of individuals, users might have less exposure to pages displaying advertisements. For several analysts, this could generate resistance from major internet platforms, especially those whose revenue depends on advertising.

Meta, for its part, signals partnerships with Shopify, Instacart, and Dick’s Sporting Goods. Muse's advancement also generated reactions in other segments: there was a decline in the stocks of asset managers, brokers, and banks, with Charles Schwab registering a drop of over 6%. Booking Holdings and Allstate were also impacted.

Meta holds a crucial advantage, given that Facebook and Instagram offer vast distribution infrastructure, in addition to their services already accumulating data from millions of users. Currently, the only comparable product is the Instinct startup's agent, which is available only by invitation.

However, this landscape is expected to change soon. Analysts predict that OpenAI will announce a consumer-focused AI agent in the coming weeks. Google is also cited as a potential competitor, and Apple may enter this market later.

In this context of fierce competition, being a pioneer can be decisive. According to Ken Gawrelski, an analyst at Wells Fargo, the effectiveness of these assistants improves as consumers become familiar with them and direct their functionalities. Furthermore, Muse demands high computational capacity, as each agent runs on its own virtual computer, which represents a factor in the dispute; however, analysts cited by the Wall Street Journal indicate that Google possesses sufficient infrastructure to develop a similar solution.

So far, the rapid implementation of Muse has positioned Meta at the epicenter of this new dispute. However, the agent's success will not depend solely on the number of users, but also on the willingness of platforms to allow an artificial intelligence to perform tasks on behalf of their customers.

Meta tests use of outsourced human agents to complement AI agents
Read more
tecnoblog.net

Meta tests use of outsourced human agents to complement AI agents

Meta implemented a test using outsourced human agents in phone calls that would otherwise be managed by the Muse Artificial Intelligence agent. This experimental feature, called 'human concierge,' allows human collaborators to discreetly take over a call in certain situations.

The concierge concept aims to complement an existing capability in Muse, which already allowed users to use the assistant to contact establishments, schedule appointments, or check product availability, among other activities. These tests had been running since August.

According to messages obtained by the Reuters agency, the company itself clarified that 'Muse is now capable of forwarding requests to a trained agent, who makes the call and handles the entire process.' However, this operation with human support was halted after receiving internal criticism.

According to the agency's investigation, the internal complaints were motivated by concerns regarding user privacy. The main apprehension lies in the possibility of personal data and other sensitive information being exposed to outsourced customer service employees.

Although there is little detailed information about the exact functioning of the tool, the term 'discreet redirection' suggests that the change of control may not be clearly signaled to the user.

In response to the negative reaction, an executive from the Superintelligence Lab acknowledged in a subsequent statement that it 'was a failure' to start the tests with operators without providing proper transparency warnings. She confirmed that the novelty has been temporarily suspended.

On the other hand, spokesperson Daniel Roberts contested the agency's leaks, stating that the feedback received was 'overwhelmingly positive' and defending the system's refinement. He declared: 'We are working with merchants to continue refining this potential calling feature and will only launch it when it is ready.'

This persistence occurs parallel to the recognition that human intervention increased the tool's effectiveness, something common in customer service. An internal document indicated that the success rates of this new service range between 95% and 98%, a higher index than achieved by calls made exclusively by artificial intelligence.

This innovation follows the highly successful launch of Muse, which took the lead in the application download ranking in the United States, reaching over 2.5 million installations in the first two weeks. It is important to note that the application is not yet available in Brazil.

The assistant was designed to operate autonomously in routines such as sending emails, online shopping, and travel bookings, competing with other tools like Gemini Spark and Claude Cowork.

During its presentation, Meta strongly emphasized security, guaranteeing that each user would have access to an isolated 'virtual machine' in the cloud to protect sessions, and that sensitive credentials would be stored in specific encrypted vaults. Amazon is among the major companies that chose to block the agent.

Popular