Agent confidence on the technical frontier

72d ago · US · primary source: technologyreview.com

A survey of 300 global technology experts ranks 101 agent tasks, revealing where organizations are confident deploying agentic AI and where human oversight remains critical, according to a new MIT Technology Review Insights report sponsored by Microsoft [1]. The report identifies 2026 as an "inflection year" for aligning AI projects with strategic business objectives, citing Gartner [1]. IT infrastructure costs are projected to grow two to three times by 2030 even as budgets stay flat, per McKinsey [1]. Tech teams are already putting agents to work to automate tasks and manage workflows [1]. Confidence is highest for structured, measurable processes such as generating reports and boilerplate code [1]. Data workflows represent the breakthrough domain, with experts trusting agents most where structure provides a reliable foundation for decisions — including data quality monitoring, visualization anomaly detection, real-time data stream monitoring, and data profiling [1]. Human oversight remains a key factor of success [1]. Jeremy Winter, corporate vice president and chief product officer at Microsoft Azure Platform, stated: "As we design agents to operate within the same operational boundaries, identity systems, and governance models that teams already use, they start to behave more like the systems organizations already trust" [1]. The push toward agentic AI unfolds against a backdrop of rapid large language model advancement. LLMs, typically based on transformer architecture, are foundational to modern chatbots and can generate, summarize, translate, and analyze text [3]. Generative pre-trained transformers are pre-trained to predict the next word and then fine-tuned to follow instructions [3]. Microsoft Copilot, launched in 2023, utilizes a model built upon OpenAI's GPT large language models and was introduced as a built-in feature for Microsoft Bing and Microsoft Edge before expanding across Windows and Microsoft 365 [7]. Confidence in agents drops when tasks require deeper business context, which remains at an early stage of development [1]. The field of AI safety, which gained significant popularity in 2023, focuses on preventing accidents, misuse, or harmful consequences from AI systems and encompasses AI alignment to ensure systems behave as intended [6]. Researchers have expressed concern that AI safety measures are not keeping pace with the rapid development of AI capabilities [6]. Artificial general intelligence — a hypothetical AI that matches or surpasses human capabilities across virtually all cognitive tasks — remains a stated goal of companies including OpenAI, Google, xAI, and Meta [5]. A 2020 survey identified 72 active AGI research and development projects across 37 countries [5]. The experts surveyed for the MIT Technology Review report expect agent confidence to accelerate as experience deepens and business environments mature [1].

application

Background sources we checked (8)
  • en.wikipedia.org ↗ U.S. Agent (John Walker) is a character appearing in American comic books published by Marvel Comics, usually those starring Captain America and the Avengers. Created by Mark Gruenwald and Paul Neary, the character first appeared in Captain America #323 (November 1986) as Super-P…
  • en.wikipedia.org ↗ A large language model (LLM) is a neural network trained on a vast amount of text for natural language processing tasks, especially language generation. LLMs can typically generate, summarize, translate, and analyze text in many contexts, and are a foundational technology behind …
  • en.wikipedia.org ↗ Ebola, also known as Ebola virus disease (EVD) and Ebola hemorrhagic fever (EHF), is a zoonotic viral hemorrhagic fever in humans and other primates, caused by four of the six known ebolaviruses. Symptoms typically start anywhere between two days and three weeks after infection. …
  • en.wikipedia.org ↗ Artificial general intelligence (AGI) is a hypothetical type of artificial intelligence that matches or surpasses human capabilities across virtually all cognitive tasks. Beyond AGI, artificial superintelligence (ASI) would outperform the best human abilities across every domain …
  • en.wikipedia.org ↗ AI safety is an interdisciplinary field focused on preventing accidents, misuse, or other harmful consequences arising from artificial intelligence systems. It encompasses AI alignment (which aims to ensure AI systems behave as intended), monitoring AI systems for risks, and enha…
  • en.wikipedia.org ↗ Microsoft Copilot is a generative artificial intelligence chatbot developed by Microsoft AI, a division of Microsoft. Based on the Microsoft Prometheus large language model, it was launched in 2023 as Microsoft's main replacement for the discontinued Cortana. The service was intr…
  • en.wikipedia.org ↗ Listas was a tool for the creation, management and sharing of lists, notes, and favorites, being developed at Microsoft Live labs. It allowed users to quickly and easily edit lists, share them with others for reading or wiki-style editing, and discover the public lists of other u…
  • en.wikipedia.org ↗ Roopam Sharma (born 24 May 1995) is an Indian scientist. He is best known for his work on Manovue, a technology which enables the visually impaired to read printed text. His research interests include Wearable Computing, Mobile Application Development, Human Centered Design, Comp…

Sources

Spot something wrong? Report an issue