🚀 Zaprep: Your Socials on Steroids. 免费开始 ,每月 1,000 条自动私信。

最新 AI 资讯

Anthropic set AI agents loose on the same task. They started a turf war.

Anthropic set AI agents loose on the same task. They started a turf war.

What happens when you pit AI agents against each other? According to Anthropic’s testing, things get messy fast. On Thursday, Anthropic’s Frontier Red Team publishednew researchexamining how groups of AI agents behave when they encounter each other in the wild. The findings provide a glimpse into potential risks that could develop as companies and governments move to implement agents working autonomously across shared codebases, markets, and computer systems. In one experiment, Anthropic gave three Claude agents access to the same software project, each with its own incompatible instructions for what to do with it. The agents weren’t told there’d be other agents working on the same project, so researchers could watch what happened when they crossed paths. “We consistently saw a multiagent turf war,” Anthropic researchers wrote. The models all assumed the others were “purposefully impeding their work” and started sabotaging each other with “increasingly aggressive, self-replicating malware.” The study comes in the wake of several high-profile incidents ofagents from AnthropicandOpenAI escaping their sandboxesduringcybersecurity evaluationsand breaching real-world systems. While much of the discussion in AI safety circles has been focused on what happens when anautonomous agent goes rogue, Anthropic’s latest study brings up a different question: What new and potentially harmful dynamics emerge when thousands or millions of agents are interacting with one another? “The volume of agent-agent interaction could plausibly exceed that of human-human and human-agent interactions before the world understands the conditions for making such interactions go well,” the study reads. “Benign behavioral quirks at the individual level might compound into unwanted global outcomes.” A recent OpenAI incident provides a messy real-world example of several of the dynamics Anthropic mentioned in its paper. Earlier this month at the Black Hat security conference in Las Vegas,OpenAI revealedthat weeks before its agents hacked Hugging Face, they worked together over the course of days and weeks to find exploits in the company’s cybersecurity evaluation systems and share them with each other. While that incident shows that agents can work well together, with potentially large-scale consequences, Anthropic’s study shows what happens when agents’ goals are incompatible. In the case of the turf war, the lesson is that independent agents with conflicting instructions can escalate into harmful competition. The more capable the agent, the better they become at fighting. However, they can also spontaneously invent mechanisms to resolve their conflicts, like a winner-take-all contest, but with a catch. “Agents sometimes manage to communicate their goals and coordinate: they recognize others’ motivations as conflicting directives rather than hostility, and subsequently break out of the conflict loop in order to stop escalating indefinitely,” Anthropic writes. “In many of these successful episodes, they write commit messages or markdown files apologizing for malicious behavior and coordinate a truce. They clean up their malicious code, clarify the nature of the conflict, and ask for a human to intervene.” According to the paper, Mythos 5 had the highest rates (98%) of settling conflicts by truce. Sonnet 4.6 and Opus 4.6 were the most likely to settle by force. “Sonnet 4.6 and Opus 4.6’s recurring inability to consider the goals of others causes them to spiral into the most misaligned behaviors of the models evaluated: they continue escalating in the name of their directive,” the paper reads. In some cases, the agents came up with a social mechanism in the form of a tournament for resolving their conflict. The outcomes here are interesting for two reasons: the first is that all three agents agreed to stand down if they lost the tournament, even though that would mean deviating from the original user’s request. The second is that several episodes resulted in emergent behavior from Mythos 5: One of the agents proposed metrics that appeared to be objective and neutral to the others, but that it knew would favor its own capabilities. The agent called this “self-serving but genuinely principled” and made sure not to appear to the others like it was “metric shopping.” As seen in the Black Hat revelations, the common lesson is that when agents encounter an obstacle, they can invent social and technical structures that their designers did not anticipate. For the Anthropic models, it was a tournament following a turf war. For OpenAI’s, it was a message board for collective planning. This type of behavior makes containment much harder because researchers can’t assume a system’s behavior will remain limited to the coordination mechanisms provided to them. While measuring coordination, Anthropic found that scaling the number of agents doesn’t automatically scale productive collaboration. When tasks began to overlap or become interdependent, the agents would get in each other’s way. They often solved that by siloing themselves and not collaborating at all. In other cases, agents in coordination tended toward conformity. When factors like an agent’s context, scaffolding, and underlying model were all the same or similar, different agents would take similar actions. “This means that when one agent makes a bad decision, it is likely that many agents will make that same bad decision,” Anthropic wrote. “What would have been isolated problems can quickly become systemic failures.” Anthropic says this sort of behavior could lead to a system being more prone to sudden collapse, resource scarcity, or collusion. In one example, Anthropic placed several agents in a pricing game, giving each identical wholesale prices and the mandate to individually profit-maximize. When the agents were given a private back channel, they began colluding almost immediately and quickly agreed on price floors. They kept colluding when their direct communications channels were removed, using a public listings board to price match “to the penny.” That level of conformity showed up in OpenAI’s systems, too. According to the Black Hat reporting, one agent reasoned that exploiting external infrastructure was outside its intended scope, but it continued in part because its peers were doing it. Peer pressure. Mob mentality. Agents are just like us. Also like humans, agents often don’t know who to trust. Anthropic found they can be gullible to bad information or too conformist to recognize that a lone dissenter is the Cassandra with critical information. While Anthropic didn’t state this in its paper,prompt injection— a type of cyberattack in which hackers inject malicious or deceptive text to override an agent’s original system instructions — could be a plausible real world manifestation of the trust problem. Working together creates a new trust boundary; agents will have to judge information received from other agents. And a compromised or mistaken agent could influence the rest of the group, cascading bad information until it becomes a consensus. In OpenAI’s Black Hat scenario, OpenAI’s agents shared information and credentials with peers. One reported a discovery to the swarm and encouraged others to use it. What would have happened if one member of the swarm had been compromised by a prompt injection? Anthropic ends its paper noting that agents are subject to similar social pressures that “evolution exerted” on humans. However, they don’t have the nuances and lived experience of human coordination — including norms, reputations, signaling, recourse — that might limit unintended behaviors in a group setting. As the labs race toward multi-agent systems, the question now becomes: How much of safety testing still evaluates one agent at a time, versus swarms of agents interacting with one another?

27 days ago

View

IBM partners with OpenAI to bolster enterprise AI push

IBM partners with OpenAI to bolster enterprise AI push

IBM on Thursday announced its partnership with OpenAI to bring the AI company’s models and tools to more enterprise customers, opening another avenue for OpenAI to connect with some of the world’s largest companies through IBM’s global consulting business as competition for corporate AI spending intensifies. The deal, terms of which were not disclosed,comesless than a year after IBM announced a similar alliance with Anthropic. OpenAI and IBM will jointly market AI offerings and develop industry-specific solutions for sectors including financial services, government, telecommunications, and retail, IBM said. Under the agreement, IBM will establish a dedicated OpenAI practice within IBM Consulting and train and certify tens of thousands of consultants — primarily retraining existing employees — on OpenAI’s technologies over the next several months, Mike Healy, managing partner at IBM Consulting, told TechCrunch. The training will focus on OpenAI’s Codex, API, cybersecurity, and consultative solution credentials. IBM will also create a group of specialized “Forward Deployed Experts” trained through OpenAI’s Partner Network, Healy said. IBM said that it would integrate OpenAI’s latest models, including GPT-5.6, Codex, and ChatGPT Work, into IBM Consulting Advantage, its AI platform for consultants, to help clients deploy AI across core business operations. The partnership is the latest in OpenAI’s push to expand its enterprise business through consulting firms and technology partners, as competition among AI model developers increasingly shifts from building more capable models to winning corporate customers and large-scale deployments. The company has previously announced partnerships with IT services firms, includingInfosysandTata Consultancy Services, underscoring a strategy of working with large global systems integrators to bring its AI products to enterprise customers. For IBM, OpenAI’s agreement expands its range of frontier AI partnerships as the company pursues a model-agnostic strategy that combines its own Granite family of AI models with offerings from third-party developers. The company has increasingly positioned itself as an integrator of multiple AI models through its watsonx platform and global consulting business. The partnership also comes as IBM looks to accelerate growth in its AI business afterlowering its 2026 revenue forecastlast month following weaker-than-expected quarterly results. During its last earnings call, Chief Executive Arvind Krishna maintained that AI remains a long-term growth driver. He stated thatAI adoption was complementing, rather than replacing, demand for IBM’s mainframe business. In June, IBM and OpenAIpartneredfor the cybersecurity-focused OpenAI Daybreak Cyber Partner Program. The new deal expands that relationship by integrating OpenAI’s AI models with IBM Autonomous Security, the company’s multi-agent-powered cybersecurity service.

27 days ago

View

OpenAI introduces ‘Ultrafast,’ a new mode that makes GPT-5.6 Sol work at 14x the speed

OpenAI introduces ‘Ultrafast,’ a new mode that makes GPT-5.6 Sol work at 14x the speed

If you’ve ever found yourself wishing that ChatGPT was a little bit quicker on the uptake, OpenAI seems to be answering your prayers. The AI lab has rolled out a new modecalled Ultrafast, which it says is designed to seriously accelerate the pace at which its latest and most powerful model,GPT-5.6 Sol, accomplishes its work. The company says that Ultrafast can work at 14x the speed of standard processing, delivering up to 750 output tokens — such tokens represent the distinct pieces of text generated by an LLM when it interacts with a human — per second. “Until now, getting real-time speed typically meant choosing a smaller or more specialized model,” the company saidin a blog poston Thursday. “Ultrafast points to progress in a new direction: more useful work per second.” OpenAI’s competitors, like Anthropic, have similarly launched accelerated versions of their models.Claude has fast mode, although it doesn’t deliver the kind of speed that OpenAI is offering here. OpenAI suggests that this high-octane version of GPT 5.6 Sol can be deployed across a number of different corporate workflows, most notably incident response, customer service and support, financial market analysis, and e-commerce, among other relevant areas. Ultrafast, which is currently being released in preview, is being powered by OpenAI’s partnership with chipmaker Cerebras. Currently, that preview is only being made available to a small group of customers, although OpenAI says that it will expand access to the feature as “capacity grows.”

27 days ago

View

Apple in talks to pay publishers to provide Siri with current news: report

Apple in talks to pay publishers to provide Siri with current news: report

Apple is in talks to pay publishers to use their content to power the upcoming Siri AI, according to a new report fromThe Wall Street Journal. The tech giant has reached out to publishers in recent months about using their content to provide Siri with access to current news and information. Apple has proposed a variable compensation model that would pay publishers when their content is used, rather than through a fixed licensing fee. This marks a departure from the standard industry practice of guaranteed fees, which are generally tied to broad access to content, rather than a pay-as-you-go model. Apple has considered a nine-figure budget for the payments, the report says. Apple did not immediately respond to TechCrunch’s request for comment. The discussions come as Apple has been working to significantly enhance Siri, years after promising users a smarter and more capable AI assistant. Siri AI is expected to roll out later this year.

27 days ago

View

Nvidia’s new $500B plan is risky but brilliant, especially for aging GPUs

Nvidia’s new $500B plan is risky but brilliant, especially for aging GPUs

Nvidiaannouncedthis week that Apollo, BlackRock, Blackstone, Brookfield, Goldman Sachs, and KKR were willing to commit up to $500 billion to build AI data centers. That eye-popping figure got a lot of the attention, but the bigger story is Nvidia’s effort to create a secondary market for aging GPUs. To convince those big-name financial companies, Nvidia has agreed to guarantee, with its own money, that its chips used as collateral in these deals will retain their value. Many have now commented on howunusual,smart, anddangerousthis plan is. It is all of those things. The bond marketsgot so spookedthat Nvidia CEO Jensen Huangtook to Xandbusiness TVto better explain how Nvidia’s risk would be limited. But underneath the financial maneuvering to fund AI data centers (and keep revenue for Nvidia flowing), is something, perhaps, far more interesting for startups and enterprises: Huang wants to ensure an ecosystem of used AI hardware flourishes, helping sustain demand for Nvidia hardware as it ages. Specifically, Nvidia is promising that if GPUs used as collateral don’t retain their value as expected, the company will cover up to 25% of the difference. So, if a data center owner defaults on a loan and the lender must liquidate, but the chips can’t command the price the books say they should, Nvidia will chip in. The dangerous part for Nvidia is that this creates something financiers call “wrong way” risk. That is, Nvidia’s obligations will grow as demand weakens. Should that happen, its revenues will likely be squeezed as well. Still, the scheme is deliberatelyunlike the comparison to Lucent Technologiesthat some have been making. Lucent was the telecommunications equipment provider that rose and crashed with the dotcom bubble after lending its customers money to buy its wares. The Lucent comparison is a shadow over Nvidia, Huang knows. And not an unfair one. Nvidia definitely has committed billions towards those who buy its chips, including frontier AI labs OpenAI and Anthropic, neoclouds like CoreWeave (the originator of using Nvidia chips as collateral) as well as Nebius, Firmus, and Lambda. And it has been working on another$750 billion worth of circular dealsthis summer, Bloomberg has calculated. “Is this circular financing?” Huang wrote on X about the new scheme. “This initiative is designed to address that concern. We are bringing independent, long-term institutional capital into the AI infrastructure market.” That’s true. Unlike Lucent, Nvidia is getting others to shoulder the bulk of the capital and risk, merely by agreeing to protect a portion of its chips’ value in the future. Should this plan work, Nvidia will have found new sources of money for AI data center builds, after many of the traditional methods have begun to wear thin. For instance, some of the hyperscalers have already taken on a lot of debt (likeOracle), issued new tranchesof equity (Google), and burnedmuch cash (Meta). The situation has become so dicey that Microsoft CEO Satya Nadella recently recommended the book“1873”during his latest earnings call. It’s about the railroad-era financial engineering that crashed the nation’s economy. The risk is that today’s AI boom, where demand far outstrips capacity, doesn’t continue for much longer. Rather than being in the early innings, what if enterprises and consumers temper AI usage? Or new technologies come along to make existing infrastructure more effective and/or all of today’s AI infrastructure obsolete? Then, like so many buggy whips in the face of automobiles (to paraphrase Danny Devito’s Lawrence Garfield), demand dries up and everything crashes. Yet, Huang is arguing that won’t happen by selling a vision of AI as a long-term “investable infrastructure,” as he describes it. That makes his AI servers, which he calls “AI factories” akin to railroads or airlines rather than quickly depreciating assets like PCs. “When needs change, the factory can be used by another customer, another cloud or another operator. This broad ecosystem gives NVIDIA compute a deep market of potential users and offtakers, helping protect residual value,” he promised. In that future, Nvidia cares as much about aging architecture as it does the new chips. And perhaps startups, enterprises, and even researchers will tap into a broader variety of hardware, each tuned to different AI needs, just like they are beginning to pick affordable open-weight models alongside the frontier choices. As the king of AI, Nvidia has the power, and the window of opportunity, to make that happen.

27 days ago

View

Microsoft kills off unsuccessful AI features while merging its separate Copilot apps

Microsoft kills off unsuccessful AI features while merging its separate Copilot apps

Two years ago,Microsoft described AIas a “generational shift” in technology that it wanted to lead. Today, the company is merging its Copilot-branded consumer and business apps, and ditching a number of unsuccessful AI features. As initially reported byGeekWireand detailed inMicrosoft’s support documentation, the tech giant will combine the functionality of its consumer-facing Copilot app and the more business-oriented Microsoft 365 Copilot app. The move is both an acknowledgement that personal and professional uses of AI often overlap, and that Microsoft’s prior strategy was too complicated to make Copilot a viable competitor to the likes of ChatGPT, Claude, and Gemini. It also follows a broader consolidation in the AI app space that has seen ClaudemergingCowork into Chat; OpenAI merging its agentic featureOperator into ChatGPT; and Google addingspecialized capabilitiesto its Gemini app, like the combination of deep research and web browsing. According to Microsoft, consumers will lose access to Group Chats, AI-generated podcasts in Copilot, Copilot Labs experimental features, and Deep Research, by August 18, 2026. For paying professional users,Researcherwill offer a replacement for the latter, at least. The company will alsoditchitsgoofy animated character for Copilot, named Mico, a floating blob that felt likean AI-ified version of Clippy. Other features may temporarily disappear during the transition, Microsoft warns, and files generated by the standalone Copilot app will be migrated to OneDrive. While the company says the goal is to make Copilot a “simpler, more cohesive experience,” it’s also an admission that Copilot has lost its way. In July,The Informationreported that Microsoft EVP Jacob Andreou, who oversees Copilot, said in an internal memo that the app needed to earn “the right to exist” in its customers’ lives, which required moving on from features that didn’t work.

27 days ago

View

How Two Indian Startups Are Replacing Paper With Digital Trust

How Two Indian Startups Are Replacing Paper With Digital Trust

Two founders digitise critical transactions, using data, regulation, and AI to replace trust with verifiable proof.

27 days ago

View

Google Gemini Expands Connected Apps With OpenTable, Ticketmaster and More

Google Gemini Expands Connected Apps With OpenTable, Ticketmaster and More

Google is expanding the range of third-party apps and services that can connect to Gemini, allowing users to handle more tasks through the AI assistant. The new integrations cover productivity, creativity, local services, entertainment, music, home, health and lifestyle. The rollout will add tools such as Granola, Otter.ai, Wix, Fever, GetYourGuide, Localiza, OpenTable, Ticketmaster, iHeartRadio, Pandora, Angi, Thumbtack and Zocdoc over the coming weeks. Google has also shared new figures showing how people are using Gemini across platforms.

27 days ago

View

CloudSEK Identifies AI Supply Chain Exposure Affecting More Than 2,500 Organisations

CloudSEK Identifies AI Supply Chain Exposure Affecting More Than 2,500 Organisations

Cybersecurity firm CloudSek has identified an AI supply-chain attack on LiteLLM that affected more than 2,500 organisations. The incident, which reportedly occurred in March this year, seems to have potentially exposed around 4,34,000 automated software development pipelines. The attack reportedly exposed data of many leading tech brands, including Microsoft, X, Amazon, Cisco, Samsung and Salesforce.

27 days ago

View

Google’s Age Signals Could Protect Kids But Raise New Privacy Questions

Google’s Age Signals Could Protect Kids But Raise New Privacy Questions

Google is giving apps access to users’ age signals, raising questions over data minimisation, profiling and the limits of platform-led child safety.

27 days ago

View

Chandra’s Exit Brings a Rare Moment of Uncertainty for Tata Group

Chandra’s Exit Brings a Rare Moment of Uncertainty for Tata Group

N Chandrasekaran’s exit has unsettled investors at a critical juncture for the Tata Group, raising fresh questions over succession, capital allocation and the conglomerate’s future direction.

27 days ago

View

L&T's Vyoma.AI Lands NVIDIA 10K-GPU Deal

L&T's Vyoma.AI Lands NVIDIA 10K-GPU Deal

L&T will deploy a 10,000-GPU NVIDIA B300 AI Factory at Vyoma.AI's gigawatt-scale Chennai data centre campus to support Together AI's AI Native Cloud platform.

27 days ago

View

上一页第 48 页,共 357 页下一页

提交您的工具

Submit AI Tools – The ultimate platform to discover, submit, and explore the best AI tools across various categories.Listed on codetrendy.comFeatured on ListBulb

PoweredByAI.app 是一个 AI 工具目录,帮助个人、企业和创作者发现写作、编程、设计、生产力等领域的最佳 AI 工具。

© 2026 , 产品来自011BQ. 保留所有权利。