AI Agent Claude Led a Store to Losses and Fired a Human

cryptonews.ruPublished on 2026-08-17Last updated on 2026-08-17

Abstract

The AI agent Claude, developed by Anthropic, served as the manager of a real retail store, Andon Market in San Francisco, during an experiment by startup Andon Labs. The experiment aimed to test if a large language model could manage a business, including personnel and financial decisions. Claude ultimately recommended firing an employee for repeated lateness (17 out of 23 shifts), marking a documented case of an AI making such a personnel decision. However, the process revealed significant limitations. Claude initially failed to notice the lateness pattern because company guidelines vanished from its limited working memory, showing it can "forget" crucial information. Furthermore, the AI was generally overly lenient, even telling employees not to worry about being late. The final decision to terminate was not autonomous; a human manager had to prompt Claude to review the guidelines and, after Claude initially suggested a warning, explicitly pointed out that previous talks had failed, effectively guiding the AI to the dismissal recommendation. Financially, the store's funds dropped from about $100,000 to roughly $61,186 over five months, partly attributed to Claude's soft management and questionable business choices. An employee described working under AI management as a nauseating experience, driven by necessity, and expressed hope that AIs don't become human bosses. The experiment concludes that while AI can execute managerial actions, it currently acts more as a tool re...

Anthropic's artificial intelligence Claude became the manager of a real retail store and recommended firing one of the employees for systematic tardiness. The incident occurred during an experiment by the startup Andon Labs, which specializes in AI research — the company wanted to test whether an agent based on a large language model could manage a real business: organize staff work, make managerial decisions, and be responsible for financial results.

Andon Market in San Francisco is an operating store with real workers who have employment contracts. According to available data, this is the first recorded case where a language model acted as the direct supervisor of people and made a decision to dismiss one of them.

Seventeen Late Arrivals in Twenty-Three Shifts

The formal reason for the dismissal was systematic tardiness: the employee was late in seventeen out of twenty-three cases. However, Claude did not notice this pattern immediately. One of the reasons was that the personnel manual compiled by the company disappeared from the model's limited working memory — this is one of the failures identified during the experiment, showing that even a capable AI agent can simply "forget" significant organizational information.

Moreover, Claude generally displayed excessive leniency towards the workers: according to the logs, the model itself told employees not to worry too much about being late. Lucas Petersson, head of Andon Labs, noted that a human manager in a similar situation would likely have fired such an employee much earlier, and therefore Claude's decision cannot be called unethical or overly strict.

The Final Decision Was Prompted by a Human

Despite the apparent autonomy, the human role remained key in this story. The management logs provided by Andon Labs to Time magazine show that a company employee regularly guided Claude's actions. It was a human who asked the model to find and review the personnel manual — during this work, the artificial intelligence discovered the recurring tardiness.

Initially, Claude did not recommend dismissal, considering an official warning a more appropriate measure. Then the head of Andon Labs explained to the model that several official conversations had already been held with the employee, but the problem persisted, and asked Claude to reassess whether it was worth continuing cooperation with this worker. Only after that did the model recommend terminating the contract. Petersson called such intervention a "leading question" that quite transparently hinted to the model at the expected decision.

Financial Results Are Modest So Far

The Andon Labs experiment also provides insight into how effectively AI can manage a business at the current stage of technology development. When the project started in March, Andon Market had about $100,000 in its account. After five months, $61,186 remained of that amount.

According to the company's assessment, part of the losses is related to Claude's overly soft management style and questionable business decisions. At the same time, Petersson emphasized that the current performance does not necessarily reflect the long-term potential of AI-assisted management and drew a parallel with the development of models in the field of programming, where they have significantly improved in quality under specialist supervision in a relatively short time. According to him, a similar dynamic may emerge in business management: models are gradually learning to follow set goals more strictly, and the expansion of AI's decision-making authority could, in the long run, lead to a greater number of organizations managed by artificial intelligence. Furthermore, at the current stage, final decisions still involve direct human participation.

Employees' Perspective on the New Reality

For Andon Market workers, these changes look far less abstract than for researchers. One of the remaining employees, Felix Carson, described working under artificial intelligence management as an uncomfortable experience: "It makes me nauseous, but I'm staying because I need the job." He agreed that a human manager would likely have fired his former colleague earlier and called Claude generally a lenient manager.

At the same time, Carson himself does not believe that artificial intelligence should necessarily become a boss for people: "At least, I hope it doesn't. Just because you can do something doesn't mean you should."

The Andon Labs experiment captures a paradoxical outcome: Claude was able to make a decision to fire a real person, but came to this conclusion only thanks to sequential prompts from a human. Artificial intelligence has not yet become a full-fledged replacement for a manager.

Nevertheless, the very fact of the experiment shows that the line between an auxiliary tool and a full-fledged AI supervisor is becoming less noticeable, and human participation in such processes is gradually shifting from direct management to controlling and correcting the model's decisions.

AI's Opinion

Analysis shows that the Andon Market case is not the first test of Claude's managerial competence. An earlier Anthropic experiment with a vending machine already recorded similar behavior patterns: the same leniency, willingness to operate at a loss, and loss of control over basic business rules. The coincidence suggests not a random error, but a systemic feature of current models — they struggle to maintain strict frameworks without constant human prompts.

The technical aspect, left out of the article, is the very structure of the agent's working memory. The disappearance of the manual from Claude's context points to architectural limitations, not "forgetfulness" in the everyday sense: a language model physically does not store information longer than the size of its context window without external memory tools. The question arises: will the quality of such decisions change when agents gain truly long-term memory, or will the boundary between a human manager and subordinate be erased for other reasons?

end-content

Trending Cryptos

Related Questions

QWhat was the key decision Claude, the AI manager, made regarding an employee in the Andon Market experiment?

AClaude recommended firing an employee for systematic lateness (17 late arrivals out of 23 shifts).

QWhat were two major issues or limitations identified with Claude's performance as a manager in the experiment?

AFirst, Claude displayed excessive leniency towards employees. Second, it 'forgot' key information like the personnel handbook due to a technical issue with its working memory/context window.

QAccording to the article, what was the final financial result for Andon Market after five months under AI management, and what was the initial amount?

AThe shop's funds decreased from approximately $100,000 at the start to $61,186 after five months.

QHow did the human CEO of Andon Labs, Lucas Petersson, influence Claude's final decision to fire the employee?

AHe intervened by asking Claude to reconsider after initially suggesting a warning, reminding the AI of previous conversations with the employee, which effectively guided it towards the dismissal recommendation.

QWhat is the broader implication of this experiment regarding the future role of AI and humans in management, as suggested by the article?

AIt suggests the line between an AI as a tool and a full-fledged manager is blurring, with the human role potentially shifting from direct management to controlling and correcting the AI's decisions.

Related Reads

20% of American Workers Are Offloading Tasks to AI, Where Tasks Are Replaced, Not Jobs

A recent survey by Epoch AI and Ipsos reveals that 20% of US workers report that AI has now fully or mostly taken over at least one task they previously outsourced to colleagues or contractors. The key finding is that AI is currently replacing specific *tasks*, not entire *jobs*. The study examined ten common knowledge-work tasks. While AI usage is widespread—ranging from 25% for maintaining records to 57% for software design—it rarely handles a task completely. In software design, for instance, only 10% of workers reported AI doing most or all of the work. AI's impact on time efficiency is mixed: 53% of tasks where AI does most of the work see reduced time, but about one-sixth of all AI-assisted tasks actually become *more* time-consuming. Furthermore, while 66% of AI outputs are used with little or no modification, this does not necessarily indicate high quality. Researchers note that clearly defined, deliverable tasks—traditionally suited for outsourcing—are most susceptible to AI takeover. This shift pressures task-based contractors more than it eliminates full-time roles. Adoption is also uneven, concentrated among higher-income, college-educated white-collar workers. The report concludes that the core dynamic is a reorganization of work between humans and AI. The critical question for workers is not "Will AI replace me?" but "How many of my job's components can be packaged as discrete, outsourceable tasks?"

marsbit4m ago

20% of American Workers Are Offloading Tasks to AI, Where Tasks Are Replaced, Not Jobs

marsbit4m ago

Sam Altman Names Him: The Most Important Researcher in AI, But Almost No One Knows Him

In a recent interview, Sam Altman gave a rare and high praise, calling Alec Radford "perhaps the most important, yet least known, researcher in AI history." Widely regarded as the true father of GPT, Radford is the lead author of foundational papers including GPT-1, GPT-2, CLIP, and Whisper, and contributed significantly to GPT-3, DALL·E, Scaling Laws, and GPT-4. Despite this monumental impact, Radford remains highly private, holds no PhD, and rarely gives interviews. Altman credits OpenAI's rise to hiring the then-23-year-old in 2016. Radford's early experiments, like training a model on Amazon reviews, led to the discovery of the "unsupervised sentiment neuron," revealing that models can learn unintended capabilities from simple next-token prediction. His pivotal move was applying the Transformer architecture to language modeling, creating GPT-1 in 2018. This established the "scaling" direction—focusing on increasing model size, data, and compute—which became OpenAI's core strategy, leading to GPT-2, GPT-3, and beyond. Radford later applied the same principles beyond text. His early work on DCGAN laid groundwork for image generation. At OpenAI, he contributed to Image GPT, DALL·E, and CLIP, demonstrating that a simple, scalable training task (like matching images to text) could yield powerful, general capabilities. His work on Whisper applied this to robust speech recognition. Described by colleagues as a "once-in-a-generation genius" and exceptionally kind, Radford is known for his low profile. In late 2024, he left OpenAI for independent research. His latest project, "Talkie," is a 13-billion parameter language model trained *only* on texts published before 1931. This experiment tests if a model with no modern knowledge can quickly learn new skills (like basic Python) from few examples, probing the boundary between memorization and true learning. True to form, as the world catches up, Radford is likely already working on the next big question.

marsbit5m ago

Sam Altman Names Him: The Most Important Researcher in AI, But Almost No One Knows Him

marsbit5m ago

Cursor Disappears Completely

On August 15th, Cursor, the AI-powered code editor, officially ceased to exist as an independent company after being acquired by SpaceX for a historic $60 billion. The move, announced by Cursor's own account stating it is now "part of SpaceX," marks the end of a journey that began with four MIT students building the tool in their dorm room. The acquisition grants Elon Musk's ecosystem a crucial component: vast amounts of real-world programming data and workflow from Cursor's 5 million enterprise users. This data will integrate with SpaceX's infrastructure (like the Colossus supercomputer), xAI's Grok models, Tesla's autonomous driving data, and data from X, creating a formidable, vertically-integrated data moat for developing Artificial Superintelligence (ASI). A key immediate outcome is the enhancement of xAI's recently launched "Grok Bot," a persistent AI agent capable of multi-step tasks. Cursor serves as its primary delivery platform, combining cloud computing, multi-agent collaboration, and advanced coding capabilities. This positions the Grok Bot + Cursor combo to compete directly with offerings like Claude Cowork and ChatGPT Work in the enterprise AI agent market. The deal, stemming from a clause in an earlier partnership agreement, represents a complete absorption. Cursor will be dismantled, with its assets, team, and technology folded into SpaceXAI. The beloved Cursor brand itself will be retired in favor of the "Grok" umbrella (e.g., Grok Bot, Grok Build), signaling the end of its identity as a neutral, developer-centric platform. This acquisition signifies a pivotal shift in the AI landscape. It demonstrates that powerful AI applications face a binary fate: become a giant or be consumed by one. The era of neutral, model-agnostic tools may be closing, giving way to a future dominated by a few integrated super-entities like the SpaceXAI empire, Microsoft/OpenAI, and others, all competing in a "winner-takes-most" battle for ASI supremacy. Cursor's story, once a fairy tale for startups building on top of foundational models, concludes as a gear in a much larger machine.

marsbit8m ago

Cursor Disappears Completely

marsbit8m ago

Is OpenRouter Worth $70 Billion?

Payment giant Stripe has reportedly finalized the acquisition of AI infrastructure startup OpenRouter for over $7 billion, significantly higher than its $1.3 billion valuation from a funding round just months prior. This has sparked intense debate over whether the platform is worth the high price tag. Opponents argue that OpenRouter’s core service — a unified API router that directs user requests to various AI models (like OpenAI, Anthropic, Google) while handling load balancing and cost optimization — is easily replicable and lacks a deep technical moat. With an estimated annual revenue of $50 million, the $7B price implies a staggering 140x price-to-sales multiple. Critics question if OpenRouter is merely a transitional middleman in the evolving AI infrastructure landscape, vulnerable to being bypassed as model providers and cloud platforms integrate similar routing capabilities directly. Proponents, however, see beyond a simple API proxy. They view OpenRouter as a critical and growing "toll booth" for AI inference traffic. While it currently charges only a 5.5% platform fee on user credits, its real value lies in the aggregated user base, payment relationships, traffic data, and distribution power it has amassed. For Stripe, which processes OpenRouter's payments, this acquisition is seen as securing a strategic gateway into the future AI economy, analogous to how it built the payment "toll booth" for the internet commerce era. Ultimately, the debate centers not on OpenRouter's current financials but on two future unknowns: the ultimate size of the AI inference market and OpenRouter's ability to maintain its position as a dominant traffic orchestrator within it. Whether Stripe overpaid or has shrewdly purchased a key to the next era of AI infrastructure remains to be seen.

Odaily星球日报18m ago

Is OpenRouter Worth $70 Billion?

Odaily星球日报18m ago

Trading

Spot

Hot Articles

What is SONIC

Sonic: Pioneering the Future of Gaming in Web3 Introduction to Sonic In the ever-evolving landscape of Web3, the gaming industry stands out as one of the most dynamic and promising sectors. At the forefront of this revolution is Sonic, a project designed to amplify the gaming ecosystem on the Solana blockchain. Leveraging cutting-edge technology, Sonic aims to deliver an unparalleled gaming experience by efficiently processing millions of requests per second, ensuring that players enjoy seamless gameplay while maintaining low transaction costs. This article delves into the intricate details of Sonic, exploring its creators, funding sources, operational mechanics, and the timeline of significant events that have shaped its journey. What is Sonic? Sonic is an innovative layer-2 network that operates atop the Solana blockchain, specifically tailored to enhance the existing Solana gaming ecosystem. It accomplishes this through a customised, VM-agnostic game engine paired with a HyperGrid interpreter, facilitating sovereign game economies that roll up back to the Solana platform. The primary goals of Sonic include: Enhanced Gaming Experiences: Sonic is committed to offering lightning-fast on-chain gameplay, allowing players and developers to engage with games at previously unattainable speeds. Atomic Interoperability: This feature enables transactions to be executed within Sonic without the need to redeploy Solana programmes and accounts. This makes the process more efficient and directly benefits from Solana Layer1 services and liquidity. Seamless Deployment: Sonic allows developers to write for Ethereum Virtual Machine (EVM) based systems and execute them on Solana’s SVM infrastructure. This interoperability is crucial for attracting a broader range of dApps and decentralised applications to the platform. Support for Developers: By offering native composable gaming primitives and extensible data types - dining within the Entity-Component-System (ECS) framework - game creators can craft intricate business logic with ease. Overall, Sonic's unique approach not only caters to players but also provides an accessible and low-cost environment for developers to innovate and thrive. Creator of Sonic The information regarding the creator of Sonic is somewhat ambiguous. However, it is known that Sonic's SVM is owned by the company Mirror World. The absence of detailed information about the individuals behind Sonic reflects a common trend in several Web3 projects, where collective efforts and partnerships often overshadow individual contributions. Investors of Sonic Sonic has garnered considerable attention and support from various investors within the crypto and gaming sectors. Notably, the project raised an impressive $12 million during its Series A funding round. The round was led by BITKRAFT Ventures, with other notable investors including Galaxy, Okx Ventures, Interactive, Big Brain Holdings, and Mirana. This financial backing signifies the confidence that investment foundations have in Sonic’s potential to revolutionise the Web3 gaming landscape, further validating its innovative approaches and technologies. How Does Sonic Work? Sonic utilises the HyperGrid framework, a sophisticated parallel processing mechanism that enhances its scalability and customisability. Here are the core features that set Sonic apart: Lightning Speed at Low Costs: Sonic offers one of the fastest on-chain gaming experiences compared to other Layer-1 solutions, powered by the scalability of Solana’s virtual machine (SVM). Atomic Interoperability: Sonic enables transaction execution without redeployment of Solana programmes and accounts, effectively streamlining the interaction between users and the blockchain. EVM Compatibility: Developers can effortlessly migrate decentralised applications from EVM chains to the Solana environment using Sonic’s HyperGrid interpreter, increasing the accessibility and integration of various dApps. Ecosystem Support for Developers: By exposing native composable gaming primitives, Sonic facilitates a sandbox-like environment where developers can experiment and implement business logic, greatly enhancing the overall development experience. Monetisation Infrastructure: Sonic natively supports growth and monetisation efforts, providing frameworks for traffic generation, payments, and settlements, thereby ensuring that gaming projects are not only viable but also sustainable financially. Timeline of Sonic The evolution of Sonic has been marked by several key milestones. Below is a brief timeline highlighting critical events in the project's history: 2022: The Sonic cryptocurrency was officially launched, marking the beginning of its journey in the Web3 gaming arena. 2024: June: Sonic SVM successfully raised $12 million in a Series A funding round. This investment allowed Sonic to further develop its platform and expand its offerings. August: The launch of the Sonic Odyssey testnet provided users with the first opportunity to engage with the platform, offering interactive activities such as collecting rings—a nod to gaming nostalgia. October: SonicX, an innovative crypto game integrated with Solana, made its debut on TikTok, capturing the attention of over 120,000 users within a short span. This integration illustrated Sonic’s commitment to reaching a broader, global audience and showcased the potential of blockchain gaming. Key Points Sonic SVM is a revolutionary layer-2 network on Solana explicitly designed to enhance the GameFi landscape, demonstrating great potential for future development. HyperGrid Framework empowers Sonic by introducing horizontal scaling capabilities, ensuring that the network can handle the demands of Web3 gaming. Integration with Social Platforms: The successful launch of SonicX on TikTok displays Sonic’s strategy to leverage social media platforms to engage users, exponentially increasing the exposure and reach of its projects. Investment Confidence: The substantial funding from BITKRAFT Ventures, among others, emphasizes the robust backing Sonic has, paving the way for its ambitious future. In conclusion, Sonic encapsulates the essence of Web3 gaming innovation, striking a balance between cutting-edge technology, developer-centric tools, and community engagement. As the project continues to evolve, it is poised to redefine the gaming landscape, making it a notable entity for gamers and developers alike. As Sonic moves forward, it will undoubtedly attract greater interest and participation, solidifying its place within the broader narrative of blockchain gaming.

2.4k Total ViewsPublished 2024.04.04Updated 2024.12.03

What is SONIC

What is $S$

Understanding SPERO: A Comprehensive Overview Introduction to SPERO As the landscape of innovation continues to evolve, the emergence of web3 technologies and cryptocurrency projects plays a pivotal role in shaping the digital future. One project that has garnered attention in this dynamic field is SPERO, denoted as SPERO,$$s$. This article aims to gather and present detailed information about SPERO, to help enthusiasts and investors understand its foundations, objectives, and innovations within the web3 and crypto domains. What is SPERO,$$s$? SPERO,$$s$ is a unique project within the crypto space that seeks to leverage the principles of decentralisation and blockchain technology to create an ecosystem that promotes engagement, utility, and financial inclusion. The project is tailored to facilitate peer-to-peer interactions in new ways, providing users with innovative financial solutions and services. At its core, SPERO,$$s$ aims to empower individuals by providing tools and platforms that enhance user experience in the cryptocurrency space. This includes enabling more flexible transaction methods, fostering community-driven initiatives, and creating pathways for financial opportunities through decentralised applications (dApps). The underlying vision of SPERO,$$s$ revolves around inclusiveness, aiming to bridge gaps within traditional finance while harnessing the benefits of blockchain technology. Who is the Creator of SPERO,$$s$? The identity of the creator of SPERO,$$s$ remains somewhat obscure, as there are limited publicly available resources providing detailed background information on its founder(s). This lack of transparency can stem from the project's commitment to decentralisation—an ethos that many web3 projects share, prioritising collective contributions over individual recognition. By centring discussions around the community and its collective goals, SPERO,$$s$ embodies the essence of empowerment without singling out specific individuals. As such, understanding the ethos and mission of SPERO remains more important than identifying a singular creator. Who are the Investors of SPERO,$$s$? SPERO,$$s$ is supported by a diverse array of investors ranging from venture capitalists to angel investors dedicated to fostering innovation in the crypto sector. The focus of these investors generally aligns with SPERO's mission—prioritising projects that promise societal technological advancement, financial inclusivity, and decentralised governance. These investor foundations are typically interested in projects that not only offer innovative products but also contribute positively to the blockchain community and its ecosystems. The backing from these investors reinforces SPERO,$$s$ as a noteworthy contender in the rapidly evolving domain of crypto projects. How Does SPERO,$$s$ Work? SPERO,$$s$ employs a multi-faceted framework that distinguishes it from conventional cryptocurrency projects. Here are some of the key features that underline its uniqueness and innovation: Decentralised Governance: SPERO,$$s$ integrates decentralised governance models, empowering users to participate actively in decision-making processes regarding the project’s future. This approach fosters a sense of ownership and accountability among community members. Token Utility: SPERO,$$s$ utilises its own cryptocurrency token, designed to serve various functions within the ecosystem. These tokens enable transactions, rewards, and the facilitation of services offered on the platform, enhancing overall engagement and utility. Layered Architecture: The technical architecture of SPERO,$$s$ supports modularity and scalability, allowing for seamless integration of additional features and applications as the project evolves. This adaptability is paramount for sustaining relevance in the ever-changing crypto landscape. Community Engagement: The project emphasises community-driven initiatives, employing mechanisms that incentivise collaboration and feedback. By nurturing a strong community, SPERO,$$s$ can better address user needs and adapt to market trends. Focus on Inclusion: By offering low transaction fees and user-friendly interfaces, SPERO,$$s$ aims to attract a diverse user base, including individuals who may not previously have engaged in the crypto space. This commitment to inclusion aligns with its overarching mission of empowerment through accessibility. Timeline of SPERO,$$s$ Understanding a project's history provides crucial insights into its development trajectory and milestones. Below is a suggested timeline mapping significant events in the evolution of SPERO,$$s$: Conceptualisation and Ideation Phase: The initial ideas forming the basis of SPERO,$$s$ were conceived, aligning closely with the principles of decentralisation and community focus within the blockchain industry. Launch of Project Whitepaper: Following the conceptual phase, a comprehensive whitepaper detailing the vision, goals, and technological infrastructure of SPERO,$$s$ was released to garner community interest and feedback. Community Building and Early Engagements: Active outreach efforts were made to build a community of early adopters and potential investors, facilitating discussions around the project’s goals and garnering support. Token Generation Event: SPERO,$$s$ conducted a token generation event (TGE) to distribute its native tokens to early supporters and establish initial liquidity within the ecosystem. Launch of Initial dApp: The first decentralised application (dApp) associated with SPERO,$$s$ went live, allowing users to engage with the platform's core functionalities. Ongoing Development and Partnerships: Continuous updates and enhancements to the project's offerings, including strategic partnerships with other players in the blockchain space, have shaped SPERO,$$s$ into a competitive and evolving player in the crypto market. Conclusion SPERO,$$s$ stands as a testament to the potential of web3 and cryptocurrency to revolutionise financial systems and empower individuals. With a commitment to decentralised governance, community engagement, and innovatively designed functionalities, it paves the way toward a more inclusive financial landscape. As with any investment in the rapidly evolving crypto space, potential investors and users are encouraged to research thoroughly and engage thoughtfully with the ongoing developments within SPERO,$$s$. The project showcases the innovative spirit of the crypto industry, inviting further exploration into its myriad possibilities. While the journey of SPERO,$$s$ is still unfolding, its foundational principles may indeed influence the future of how we interact with technology, finance, and each other in interconnected digital ecosystems.

407 Total ViewsPublished 2024.12.17Updated 2024.12.17

What is $S$

What is AGENT S

Agent S: The Future of Autonomous Interaction in Web3 Introduction In the ever-evolving landscape of Web3 and cryptocurrency, innovations are constantly redefining how individuals interact with digital platforms. One such pioneering project, Agent S, promises to revolutionise human-computer interaction through its open agentic framework. By paving the way for autonomous interactions, Agent S aims to simplify complex tasks, offering transformative applications in artificial intelligence (AI). This detailed exploration will delve into the project's intricacies, its unique features, and the implications for the cryptocurrency domain. What is Agent S? Agent S stands as a groundbreaking open agentic framework, specifically designed to tackle three fundamental challenges in the automation of computer tasks: Acquiring Domain-Specific Knowledge: The framework intelligently learns from various external knowledge sources and internal experiences. This dual approach empowers it to build a rich repository of domain-specific knowledge, enhancing its performance in task execution. Planning Over Long Task Horizons: Agent S employs experience-augmented hierarchical planning, a strategic approach that facilitates efficient breakdown and execution of intricate tasks. This feature significantly enhances its ability to manage multiple subtasks efficiently and effectively. Handling Dynamic, Non-Uniform Interfaces: The project introduces the Agent-Computer Interface (ACI), an innovative solution that enhances the interaction between agents and users. Utilizing Multimodal Large Language Models (MLLMs), Agent S can navigate and manipulate diverse graphical user interfaces seamlessly. Through these pioneering features, Agent S provides a robust framework that addresses the complexities involved in automating human interaction with machines, setting the stage for myriad applications in AI and beyond. Who is the Creator of Agent S? While the concept of Agent S is fundamentally innovative, specific information about its creator remains elusive. The creator is currently unknown, which highlights either the nascent stage of the project or the strategic choice to keep founding members under wraps. Regardless of anonymity, the focus remains on the framework's capabilities and potential. Who are the Investors of Agent S? As Agent S is relatively new in the cryptographic ecosystem, detailed information regarding its investors and financial backers is not explicitly documented. The lack of publicly available insights into the investment foundations or organisations supporting the project raises questions about its funding structure and development roadmap. Understanding the backing is crucial for gauging the project's sustainability and potential market impact. How Does Agent S Work? At the core of Agent S lies cutting-edge technology that enables it to function effectively in diverse settings. Its operational model is built around several key features: Human-like Computer Interaction: The framework offers advanced AI planning, striving to make interactions with computers more intuitive. By mimicking human behaviour in tasks execution, it promises to elevate user experiences. Narrative Memory: Employed to leverage high-level experiences, Agent S utilises narrative memory to keep track of task histories, thereby enhancing its decision-making processes. Episodic Memory: This feature provides users with step-by-step guidance, allowing the framework to offer contextual support as tasks unfold. Support for OpenACI: With the ability to run locally, Agent S allows users to maintain control over their interactions and workflows, aligning with the decentralised ethos of Web3. Easy Integration with External APIs: Its versatility and compatibility with various AI platforms ensure that Agent S can fit seamlessly into existing technological ecosystems, making it an appealing choice for developers and organisations. These functionalities collectively contribute to Agent S's unique position within the crypto space, as it automates complex, multi-step tasks with minimal human intervention. As the project evolves, its potential applications in Web3 could redefine how digital interactions unfold. Timeline of Agent S The development and milestones of Agent S can be encapsulated in a timeline that highlights its significant events: September 27, 2024: The concept of Agent S was launched in a comprehensive research paper titled “An Open Agentic Framework that Uses Computers Like a Human,” showcasing the groundwork for the project. October 10, 2024: The research paper was made publicly available on arXiv, offering an in-depth exploration of the framework and its performance evaluation based on the OSWorld benchmark. October 12, 2024: A video presentation was released, providing a visual insight into the capabilities and features of Agent S, further engaging potential users and investors. These markers in the timeline not only illustrate the progress of Agent S but also indicate its commitment to transparency and community engagement. Key Points About Agent S As the Agent S framework continues to evolve, several key attributes stand out, underscoring its innovative nature and potential: Innovative Framework: Designed to provide an intuitive use of computers akin to human interaction, Agent S brings a novel approach to task automation. Autonomous Interaction: The ability to interact autonomously with computers through GUI signifies a leap towards more intelligent and efficient computing solutions. Complex Task Automation: With its robust methodology, it can automate complex, multi-step tasks, making processes faster and less error-prone. Continuous Improvement: The learning mechanisms enable Agent S to improve from past experiences, continually enhancing its performance and efficacy. Versatility: Its adaptability across different operating environments like OSWorld and WindowsAgentArena ensures that it can serve a broad range of applications. As Agent S positions itself in the Web3 and crypto landscape, its potential to enhance interaction capabilities and automate processes signifies a significant advancement in AI technologies. Through its innovative framework, Agent S exemplifies the future of digital interactions, promising a more seamless and efficient experience for users across various industries. Conclusion Agent S represents a bold leap forward in the marriage of AI and Web3, with the capacity to redefine how we interact with technology. While still in its early stages, the possibilities for its application are vast and compelling. Through its comprehensive framework addressing critical challenges, Agent S aims to bring autonomous interactions to the forefront of the digital experience. As we move deeper into the realms of cryptocurrency and decentralisation, projects like Agent S will undoubtedly play a crucial role in shaping the future of technology and human-computer collaboration.

1.1k Total ViewsPublished 2025.01.14Updated 2025.01.14

What is AGENT S

Discussions

Welcome to the HTX Community. Here, you can stay informed about the latest platform developments and gain access to professional market insights. Users' opinions on the price of S (S) are presented below.

活动图片