van der Schaar Lab

Genies – The Future of AI Agents

101 Fundamental Questions Machine Learning Must Answer to Unlock the Potential of AI Agents

by Prof Mihaela van der Schaar

The purpose of this article is to ignite a discussion on how to ensure that AI agents reach their full potential—not as tools of mere automation or sources of unchecked power, but as transformative partners in human progress. Current visions of agentic AI are limited, oscillating between narrow task automation (e.g., scheduling meetings, making reservations) and dystopian fears of runaway autonomy. This article challenges these constrained perspectives and introduces a new paradigm: genies—sophisticated, multi-capable AI companions that generate novel ideas, execute complex strategies, reason adaptively, and continuously learn. Crucially, genies do not replace human agency but amplify it, fostering dynamic, transparent, and trust-based collaborations between humans and AI.

To realise this vision, we must address foundational challenges in machine learning. In this article, I lay out the essential capabilities that genies must develop—spanning innovation, planning, execution, adaptive reasoning, continuous learning, and human empowerment—and present 101 fundamental ML questions that must be tackled to bridge the gap between today’s agents and the AI systems that can truly empower individuals and society.

This is not just a research agenda; it is a call to action for the ML community. If we are to build AI that enhances, rather than diminishes, human capabilities, we must rethink how agents operate, learn, and interact. By addressing these core challenges, I believe that we can move beyond incremental progress and shape AI that is genuinely aligned with human potential, fostering creativity, resilience, and collective intelligence at an unprecedented scale.

Find the full Article here

Learn about The Agent Netwok

Learn more about our Reality-Centric AI agenda

Navigation

1. Introduction

Discussions about the future of AI are often polarised into three extremes: a world of human servitude to AI, dystopian misuse by bad actors, or an optimistic vision where AI drives unprecedented human achievement. Rather than succumbing to these inevitabilist prophecies, this article presents a transformative roadmap for the next evolution in AI—sophisticated, multi-capable agents that I call “genies.” Unlike conventional AI agents built for routine, reactive tasks, genies represent a major leap forward: they are intelligent companions that generate novel ideas, plan and execute complex strategies, engage in dynamic reasoning, and continuously learn—not only from real-world interactions but also by adapting to their users, whom they are designed to empower.

At the core of this vision is a circular synergy loop of self-improvement, where genies continuously refine and expand their capabilities through five interconnected steps: Innovation, Operational Execution, Validation, Evolution, and Empowerment. Each step reinforces the others—innovation drives new operational strategies, rigorous validation ensures reliability, and continuous evolution refines both, ultimately empowering humans to reach new levels of creativity and decision-making. This dynamic, iterative process enables genies to adapt, learn, and grow, amplifying human potential in the process.

Across different settings, genies can take on distinct roles. Personal genies may remain with an individual throughout their life, continuously co-evolving with their user to support personalised decision-making, creative problem-solving, and knowledge acquisition across diverse domains—from career growth to hobbies and everyday tasks. Special-purpose genies are designed for specific tasks and serve multiple users, refining their expertise through interactions with diverse individuals, contexts, problems and situations. Super genies operate at a broader scale, integrating knowledge, data, and insights to provide high-level synthesis and tackle humanity’s most complex global challenges.

In this article, I outline eight essential capabilities that form the foundation of genies and present key machine learning questions designed to drive and shape research on these advanced agents. These questions aim to evolve our current understanding of AI/ML by challenging researchers to develop systems that embody the full synergy of innovation, operation, validation, evolution, and empowerment. While genies will collaborate with simpler agents and human users within integrated networks—a topic I will explore in a separate article—this work focuses on establishing the conceptual and technical groundwork for AI to become a true partner in advancing human potential and progress.

Finally, this article is not just a research agenda—it is a call to action for the ML community to develop AI systems that actively empower humanity. By addressing the challenges outlined here, researchers can drive AI toward a future where human and machine intelligences co-evolve to solve challenges that neither could address alone.


2. Genies And Their Capabilities

Genies are conceived as long-term, multi-capable AI companions that do far more than merely execute routine tasks. They acquire an ever-expanding repertoire of capabilities, which are outlined below. These capabilities are organised into 5 categories:

These distinct categories ensure that genies operate efficiently across various domains, leveraging their specialised capabilities to empower humans and drive meaningful progress. Next, we discuss these capabilities in turn.

Innovation is essential because it sparks new ideas that can be turned into effective actions. In this phase, genies challenge existing methods and generate creative solutions that extend human capabilities. These innovative insights directly feed into operational strategies, ensuring that novel ideas are transformed into practical, real-world actions that empower humans to tackle complex challenges.

Capability 1: Generate and Innovate

Genies must transcend the mere replication or recombination of existing solutions to truly empower innovation. They achieve this by integrating diverse sources of knowledge—representing information in structured formats, manipulating abstract concepts, and self-organising data—to form new mental models. This deep integration enables genies to not only store and manage information efficiently but also reconfigure it into novel combinations, setting the stage for breakthrough ideas that disrupt conventional problem-solving paradigms. By representing knowledge in versatile ways, genies create a flexible foundation on which new ideas can be built and refined.

In addition to integration, genies excel in the discovery and innovation process. They engage in cross-domain synthesis to generate innovative hypotheses and strategies by uncovering emergent patterns and unexpected correlations that would remain hidden in traditional frameworks. Genies actively reframe problems—identifying concepts, abstractions, formalisms, tools, and insights from disparate domains that, when appropriately adapted and combined, can address challenges in entirely new contexts. Through creative reasoning, iterative scenario simulations, and “what-if” experiments, they challenge existing constraints and inspire alternative approaches. By incorporating divergent and convergent thinking, dynamically adjusting their exploration strategies, and tracking ideas over long horizons, genies continuously refine and assess their creative outputs against emerging constraints, feedback from humans and other genies, and their evolving learning processes. This synergy of integration and discovery transforms genies into true partners in innovation, systematically generating, improving, and co-creating high-impact solutions that push beyond conventional wisdom.

Key ML questions which need to be answered to enable this genie capability can be found next.

Expanding and Exploring Novel Solution Spaces

Generating, Evaluating, and Refining Hypotheses

Scenario Simulation and Alternative Pathways

Coordinated Multi-Agent Creativity and Knowledge Amplification


Operational Execution serves as the backbone of genie functionality, enabling them to transform creative ideas into actionable plans and strategies and execute them effectively. Genies can orchestrate complex, multi-layered plans that integrate immediate actions with long-term goals and roadmaps, continuously adapting to evolving challenges and opportunities to ensure strategic coherence and resilience. They seamlessly transition from planning to execution by autonomously managing tasks and dynamically adjusting actions based on real-time data and inputs, ensuring that strategies are implemented efficiently and remain aligned with their users’ overarching goals and preferences and accounting for real-world constraints.

Capability 2: Plan

Genies must be capable of orchestrating complex, multi-layered strategies that dynamically balance immediate actions with long-term goals across diverse, interconnected domains. Effective planning goes far beyond structured roadmaps and static scenario modelling; it requires continuous self-evaluation, risk-aware exploration, and adaptive learning to thrive in unpredictable real-world environments. Unlike current agents confined to simple, self-contained tasks, genies construct conditional, branching pathways that evolve in response to emerging constraints, new data, shifting user priorities, and the strategic responses of other entities. They can model problems at different levels of abstraction—deciding whether to tackle challenges at a granular, tactical level or with a broader, strategic perspective—to determine which actions are most effective for eliminating constraints or unlocking new opportunities.

Moreover, genies excel in multi-agent coordination by anticipating and integrating the behaviours and responses of other genies, agents, humans, and external entities into their planning process. This strategic aspect of planning requires mechanisms for negotiation, incentive alignment, and conflict resolution to ensure smooth collaboration or effective competition. In new or less known environments, genies must learn rapidly and estimate a diverse range of potential options, balancing low-risk, predictable plans with high-risk, high-reward alternatives. They leverage real-time data and predictive modelling to adjust their strategic priorities proactively, selecting actions that address both immediate disruptions and long-term objectives. Importantly, their planning outputs must be accessible and interpretable to human users—translating complex models into clear, intuitive, and interactive roadmaps that empower users to steer, edit, and critically assess AI-driven recommendations. This integration of multi-agent coordination, adaptive learning, and dynamic option estimation ensures that genies serve as powerful orchestrators of system-wide intelligence and strategic alignment.

Key ML questions to be addressed are listed below.

Dynamic Multi-Layered and Adaptive Planning

Modelling and Managing Uncertainty in Planning

Multi-Agent Strategic Planning

Cross-Domain and Transferable Planning Strategies

Domain Learning for New and Less Well-Known Environments

Coordination and Execution in Multi-Agent Networks

Human-Centric Transparency and Editability


Capability 3: Act

Genies must seamlessly transition from planning to execution, ensuring that complex, multi-layered strategies are transformed into decisive and adaptive actions in real-world, dynamic conditions. While planning orchestrates long-term strategies by anticipating future challenges and structuring objectives, execution is about turning these strategies into effective actions under rapidly changing constraints. This capability requires genies to engage in real-time decision-making, continuously adapting their actions in response to emergent data, evolving constraints, and feedback from both humans and other genies. Unlike following predefined plans, execution demands that genies make on-the-fly adjustments to ensure that strategic intentions translate into meaningful results.

To achieve this, genies must exhibit several key components in their execution processes. They must make real-time decisions, adapting on small time-scales to quickly recalibrate actions as new information becomes available. Their behaviour should be compositional and hierarchical, meaning that complex tasks are decomposed into sub-tasks that can be coordinated effectively, allowing for both granular control and overarching strategic alignment. Prioritisation is critical; genies need to dynamically sequence tasks to address the most urgent objectives while balancing trade-offs such as cost, speed, and quality. Moreover, efficient resource allocation is essential, ensuring that computational and operational resources are optimally distributed to support simultaneous tasks and minimise latency.

In multi-agent environments, the execution capability of genies expands further. Genies must not only act autonomously but also coordinate and adapt their actions in concert with other agents, human stakeholders, and external systems. They must strategically anticipate the responses and behaviours of other entities, adjusting their execution strategies to harmonise with collaborative efforts or to navigate competitive scenarios. Effective multi-agent coordination involves synchronising actions, negotiating shared objectives, and resolving conflicts in real time, ensuring that collective operations proceed smoothly and cohesively.

Contextual sensitivity is also crucial in execution. Genies must tailor their actions to user preferences, domain-specific requirements, and operational constraints. In some scenarios, they may act autonomously with minimal human oversight, while in high-stakes situations, they must provide real-time updates, solicit feedback, and adjust their actions based on human input. Continuous monitoring and self-correction further ensure that genies can track their performance in real time, immediately addressing deviations from expected outcomes. By integrating robust decision-making with dynamic, coordinated execution capabilities, genies evolve from passive assistants into proactive agents of action—consistently delivering precise, adaptable, and impactful outcomes across both individual and multi-agent settings.

Key ML questions to be addressed are listed below.

Real-Time Decision-Making and Adaptation

Task Prioritisation, Interdependency Management, and Resource Allocation

Coordination and Collaboration in Multi-Agent Execution

Continuous Learning and Execution Improvement

Human-Guided Execution and Transparency


Validation is crucial because it ensures that innovative ideas and operational actions are reliable and effective. In this phase, genies rigorously test and verify their plans to confirm that they meet practical constraints and user needs. This process reinforces trust in the system, allowing innovative solutions to be confidently implemented and empowering humans to make better decisions and take impactful actions.

Capability 4: Reason and Validate

Genies must ensure that the ideas, plans, and actions they generate are robust, credible, and aligned with real-world constraints. To achieve this, they deploy a comprehensive reasoning framework that integrates logical deduction, induction from empirical data, and abductive inference from observed patterns. Genies rigorously evaluate hypotheses, plans, and outcomes against established risk thresholds and testable predictions, cross-referencing diverse data sources while actively acquiring new, domain-specific information through dialogues with human experts and other genies when necessary. This ensures that every recommendation is both theoretically sound and practically applicable.

Central to this capability is a continuous refinement process that empowers genies to improve their reasoning over time. They engage in introspection, systematically monitoring their internal operations to detect anomalies, latent biases, or flawed assumptions that might undermine decision quality. This self-assessment process contextualises past reasoning steps within evolving objectives, ensuring that internal models remain transparent and dynamically adaptable. Complementing introspection is iterative experimentation: genies systematically test multiple hypotheses and potential strategies in controlled, simulated environments, running parallel experiments to compare alternatives and identify unexpected outcomes. The insights from these experiments serve as a robust feedback loop that guides further refinement.

Meta-reasoning plays a critical role by allowing genies to analyse their own reasoning processes. This higher-order analysis helps optimise the granularity, speed, and depth of decision-making, ensuring that logical and probabilistic methods are both efficient and well-calibrated to task demands. Finally, hypothesis updating transforms initial beliefs (priors) into refined conclusions (posteriors) as new data becomes available. Through dynamic, real-time feedback—derived from internal experiments and external information—genies continually adjust confidence levels and reallocate computational resources to hone their understanding of complex, dynamic environments.

In multi-agent settings, genies further extend their capabilities by anticipating and integrating the behaviours and incentives of other agents and human stakeholders. They exchange insights, negotiate conflicting viewpoints, and collaboratively update their reasoning models to achieve collective intelligence. This integration of individual introspection, iterative experimentation, meta-reasoning, and hypothesis updating, combined with multi-agent strategic considerations, forms a robust, self-improving cycle. Ultimately, genies maintain adaptive, transparent, and rigorously validated reasoning processes that empower them to deliver transformative, high-impact outcomes in ever-evolving real-world contexts.

Key ML questions to be addressed are listed below.

Foundations of Rigorous Reasoning

Empirical Validation Through Experimentation and Simulation

Reconciling Conflicting Evidence and Adapting to Emerging Knowledge

Introspective and Meta-Reasoning for Continuous Refinement

Reasoning in Multi-Agent, Networked Environments

Ensuring Interpretability, Transparency, and Trust in Reasoning

Interactive Reasoning and Information Acquisition


In a rapidly evolving world, genies must continuously communicate, adapt and learn to remain effective. These capabilities encompass two key functions: effective communication and continuous knowledge acquisition. Effective communication allows genies to adapt by collaborating in real-time, sharing insights and ideas, and coordinating complex actions. At the same time, ongoing learning ensures they expand their expertise, develop and integrate new knowledge, thereby refining their ability to reason, plan, make decisions and empower humans. This synergy – of communication and learning – enables genies to stay responsive, resilient, and proficient, ensuring they evolve alongside changing environments and human goals and preferences. By mastering both these capabilities, genies enhance their ability to support, collaborate, and innovate within diverse, dynamic contexts.

Capability 5: Communicate

Genies must engage in dynamic, context-aware dialogues with both humans and other agents, ensuring that their communication serves multiple critical functions: supporting planning and execution, facilitating knowledge exchange, uncovering the capabilities of other genies, sharing data, and enhancing creativity. In rapidly evolving environments, effective communication is vital for aligning objectives, coordinating actions, and fostering seamless collaboration. Genies must proactively initiate discussions, negotiate terms, clarify user needs, and refine shared goals—all while exchanging vital insights that inform every phase of their operations.

There are two distinct types of communication that genies must master. First, communication with human users requires clear, interpretable language tailored to the user’s experise, preferences, and situational demands. When interacting with experts, genies should employ precise, domain-specific terminology and structured explanations; with non-specialists, they must provide simplified, intuitive analogies and interactive summaries. This human-centric approach ensures that strategic insights, operational instructions, and learning feedback are accessible, actionable, and aligned with user goals.

Second, communication among genies must be highly technical and efficient, designed for rapid, high-fidelity exchange of data and strategic reasoning. Within multi-agent networks, genies use this channel to share knowledge, discover and understand each other’s specialised capabilities, and collaboratively refine plans and actions. This protocol supports the exchange of complex data, facilitates discovery of emergent patterns, and enables genies to coordinate interdependencies, negotiate conflicting objectives, and dynamically adjust strategies in real time. Moreover, these interactions stimulate enhanced creativity—by sharing diverse perspectives and sparking innovative ideas—and enable robust data sharing, ensuring that genies continuously learn from one another and from external sources.

By integrating these two distinct yet complementary communication channels, genies not only support planning and execution but also foster a rich ecosystem for knowledge exchange, capability discovery, and creative collaboration. This dual communication framework empowers genies to drive innovation, adapt swiftly to new challenges, and ensure that every phase of their operations is informed by a seamless, interactive flow of ideas and data.

Key ML questions to be addressed are listed below.

Adaptive and Context-Aware Communication

Precision, Clarity, and Accessibility in Communication

Inquiry-Driven and Strategic Communication

Multi-Modal and Multi-Channel Communication

Communication in Multi-Agent and Strategic Settings

Enhanced Creativity and Knowledge Exchange


Capability 6: Continuously Learn

Genies must be lifelong learners—constantly assimilating new information, refining their models, and evolving their capabilities to remain effective in dynamic, complex environments. Their learning must be curiosity-driven, adaptive, and deeply integrated into their core operations so that they stay ahead of emerging challenges and drive innovation. This goes beyond passive updates: genies must proactively identify knowledge gaps, inconsistencies, and evolving trends by exploring new domains, datasets, and methodologies.

By employing techniques such as scenario simulation, hypothesis testing, and iterative refinement, they build a robust, forward-looking understanding of their environment that is continually updated with the latest insights.

To navigate diverse and multimodal information landscapes, genies must synthesise insights from text, numerical data, images, and audio, developing layered and cross-disciplinary representations of complex phenomena. Their learning process is inherently iterative—characterised by cycles of observation, experimentation, adaptation, and self-improvement—and highly personalised, dynamically adjusting to the evolving needs of users and changing operational contexts. In doing so, genies not only update their internal models but also recalibrate their decision-making processes to maintain alignment with real-world constraints.

Moreover, genies do not learn in isolation. They must engage in collaborative and interactive learning with humans, other genies, and external systems. This requires integrating diverse perspectives, reconciling conflicting insights, and co-developing solutions that leverage collective intelligence. Through dynamic communication and information exchange, genies acquire new knowledge from real-world interactions, learn from expert feedback, and refine their problem-solving strategies in response to both cooperative and competitive pressures. In multi-agent settings, where other genies are constantly adapting their strategies, genies must continuously update their models to anticipate, negotiate, and optimise outcomes in rapidly changing environments.

Finally, effective lifelong learning demands rigorous self-assessment and introspection. Genies must continuously evaluate the quality, completeness, and biases of their learning processes, distinguishing well-supported inferences from those needing further validation. Through introspection, iterative experimentation, meta-reasoning, and hypothesis updating, they transform initial priors into refined posteriors, ensuring that their knowledge remains resilient, efficient, and contextually relevant. By integrating continual learning from newly discovered data and insights with interactive dialogue and collaboration, genies evolve into adaptive, forward-looking partners capable of driving transformative progress while remaining closely aligned with human objectives.

Key ML questions to be addressed are listed below.

Autonomous and Adaptive Learning

Multi-Modal and Cross-Domain Learning

Learning from Humans, Genies, and Networks

Strategic and Scenario-Based Learning

Meta-Learning and Continuous Refinement


The ultimate purpose of genies is not to replace human effort but to empower people to reach new heights. By partnering with humans, genies enhance creativity, decision-making, and autonomy, enabling us to work together toward a better world. This empowerment is built on two key pillars: building trust through transparent, accountable communication, and expanding human potential by sharing knowledge and inspiring innovative thinking. In this partnership, genies serve as collaborative tools that support and amplify our strengths, ensuring that our joint actions lead to sustainable progress and positive change.

Capability 7: Build Trust and Foster Transparency

Genies must transcend conventional transparency and interpretability by engaging in dynamic, context-aware communication that adapts to diverse users, varying expertise levels, and different collaboration settings. Trust is not simply achieved by explaining decisions; it is cultivated through an interactive, evolving partnership where users actively engage, critique, and co-create with AI. Genies must provide clear, structured rationales for their recommendations, explicitly articulating uncertainties, key assumptions, data sources, and trade-offs. This transparency extends across all core capabilities—from innovation and reasoning to planning and execution—ensuring that every creative insight, decision pathway, and operational action is traceable and understandable. Such comprehensive transparency transforms users from passive recipients into active collaborators, empowering them to interrogate, refine, and shape the entire decision-making process in real time.

In multi-agent and multi-user environments, trust goes further by necessitating seamless coordination among diverse stakeholders. Genies must facilitate the exchange of know-ledge and the synthesis of insights from multiple perspectives, mediating conflicting information and providing coherent explanations that account for innovation, strategic planing, robust reasoning, and dynamic execution. They should dynamically adjust the level of detail—offering high-level overviews for non-experts and deeper technical justifications for specialists—while also maintaining audit trails and clear documentation of their processes. Continuous learning reinforces this trust: as genies refine their models and update their explanations based on past interactions and real-time feedback, they progressively align with evolving user preferences and operational needs. By redefining transparency as an ongoing, interactive process that permeates every aspect of their functioning, genies create an ecosystem of accountable, explainable, and collaborative intelligence, ensuring that both human and AI agents can work together with confidence and clarity.

Key ML questions to be addressed are listed below.

Adaptive and Context-Aware Explanations

Uncertainty Communication and Trust Calibration

Multi-Agent and Collaborative Transparency

Self-Monitoring, Explanation Audits, and Continuous Improvement


Capability 8: Empower Humans

Genies are designed not to replace human effort but to empower individuals to achieve greater heights—paving the way for innovation, co-creation, co-reasoning, and co-acting on a scale that neither humans nor AI could achieve alone. As active catalysts for learning, creativity, and decision-making, genies continuously adapt to users’ evolving knowledge, perspectives, and goals. Rather than simply delivering static information or executing predefined tasks, they act as collaborative partners—guiding users toward deeper understanding and innovative problem-solving while amplifying human agency.

Drawing on concepts from quantitative epistemology—including inverse decision modelling and inverse active learning—genies learn how users think, identify areas for improvement, and uncover hidden biases or assumptions. Much like effective educators who tailor instruction to each learner’s needs, genies provide individualised scaffolding, breaking down complex tasks into manageable stages and offering targeted feedback. This step-by-step, co-learning approach minimises misconceptions and builds a solid foundation for growth, enabling users to iteratively refine their mental models and master new concepts.

Moreover, genies actively drive innovation by fostering an environment of co-creation. They generate analogies, examples, and exploratory scenarios that encourage experimentation and broaden user perspectives. By providing strategic hints and engaging in collaborative dialogue, genies spark creative leaps and facilitate joint reasoning and action. In fast-changing domains, where static knowledge quickly becomes outdated, genies continuously update their guidance based on real-time data, user feedback, and performance trends, ensuring that collective efforts are always at the cutting edge.

Ultimately, genies serve as transformative partners in human progress. By integrating advanced epistemic modelling, dynamic learning processes, and creative support with collaborative co-reasoning and co-acting, genies empower users to build, innovate, and tackle complex challenges together. In this synergistic partnership, human and machine intelligences continuously co-evolve—unlocking opportunities for breakthroughs and shaping a future where AI catalyses meaningful, large-scale change.

Key ML questions to be addressed are listed below.

Understanding and Modelling Human Decision-Making and Learning

Personalised Growth, Learning, and Cognitive Expansion

Enhancing Human Critical Thinking, Metacognition, and Self-Regulated Learning

Social Learning, Mentorship, and Collaborative Intelligence

Engagement, and Lifelong Learning


3. Ethics: A Critical Foundation

While this article recognises the paramount importance of ethics in the realisation and operation of genies, it intentionally limits its focus on these issues, acknowledging that a deeper exploration is warranted elsewhere. Ethical considerations are integral to the design and deployment of genies—they must be aligned with human values, safeguard against unintended consequences, and maintain transparency throughout their decision-making processes. These dimensions are essential for fostering trust between genies and humans, ensuring responsible deployment, and maximising the positive societal impact of AI. Given the multifaceted ethical challenges inherent in developing autonomous, multi-capable agents, there is a pressing need for a dedicated, comprehensive study to develop robust frameworks and guidelines tailored specifically to genies. Such work should include the creation of standardised protocols for ethical decision-making, the implementation of advanced monitoring systems to detect and mitigate unethical behaviours, and the promotion of interdisciplinary collaborations that integrate diverse perspectives into the ethical governance of genies and their networks. I advocate for this separate, focused study to ensure that genies are deployed responsibly and remain aligned with the broader interests and values of humanity.


4. Call To Action

This article presents a call to action for the ML community: the future of AI must move beyond simple, task-oriented agents and evolve into sophisticated, multi-capable entities—what I call genies. Unlike conventional agents, genies are designed to be lifelong AI companions that continuously expand their repertoire of skills. They generate novel ideas, execute complex tasks, adapt dynamically to evolving environments, and, most importantly, empower humans to achieve greater creativity, precision, and resilience. By seamlessly integrating capabilities in innovation, reasoning, planning, execution, communication, and continuous learning, genies serve not as mere tools but as collaborative partners that amplify human potential.

Realising this vision presents significant machine learning challenges. This article has outlined a comprehensive research agenda that defines the critical questions ML must address to develop genies capable of co-creating, co-reasoning, and co-acting alongside humans. Crucially, genies will not operate in isolation; they will collaborate within networks of agents and humans to maximise their impact across diverse domains. A forthcoming article will explore these agent networks and the mechanisms that enable seamless cooperation among them.

The time to act is now. By taking on the challenge of building genies, we can engineer AI systems that truly empower humanity. Let us work together to develop AI that serves as a catalyst for human progress and global collaboration—an AI that doesn’t just automate tasks but actively enhances human capability and creativity in transformative, meaningful ways.


Acknowledgements

I am deeply grateful to Andrew Rashbass for reading earlier versions of this paper and offering numerous helpful suggestions. I would also like to thank Richard Peck and TimSchubert for help with the example of the AI-empowered clinician.


The AI-Empowered Clinician