REVIEW 4 major objections 4 minor 160 references
Humans Coexist, So Must Embodied Artificial Agents
T0 review · 4 major / 4 minor · reviewed 2026-08-08 · deepseek-v4-flash
Pith's one-line read The paper argues that embodied AI agents must be designed to coexist with humans through meaningful reciprocal interaction, not merely to perform tasks.
desk verdict Useful design orientation, but the 'prerequisite' claim is overreach and the formal definition rests on an unoperationalized counterfactual quality function. read the letter →
The pith
A machine-rendered reading of the paper's core claim, the machinery that carries it, and where it could break.
The reading
What carries the argument
The load-bearing mechanism is the formal definition of coexistence built around the quality function $Q_O(t)$ and the counterfactual comparison in Eqs. (3)-(4). The paper distinguishes unilateral interaction $X_t \to Y_t$ (one party's next state depends on the other's, but not vice versa) from reciprocal interaction $X_t \leftrightarrow Y_t$ (mutual dependence), and defines meaningful interaction as one that, for all observers $O$ and all times beyond a system-dependent horizon $T_S$, leaves quality no lower than no interaction: $Q_O(t' \mid X_t \to Y_t) \geq Q_O(t' \mid \emptyset)$. A coexisting agent $A^*$ is one that maintains reciprocal, meaningful interactions with the human and environment, formalized in Eq. (4). The two properties, situatedness (Eq. 5) and mutability (Eq. 6), are direct conditions on this quality comparison and carry the argument from definition to design prescription.
What would settle it
Run a matched long-term field study in which a home robot that the paper would classify as coexisting is removed for a period longer than the system's horizon $T_S$ while other conditions are held constant, and measure a proxy quality metric (voluntary user-initiated interactions, task fluency, or self-reported trust) before and during removal; if the metric does not fall after removal, the defining inequality $Q_O(t' \mid A^*) \geq Q_O(t' \mid \emptyset)$ is violated for that agent.
Extended reading notes
Core claim
The paper's central claim is that an embodied artificial agent coexists in a system only if it sustains meaningful and reciprocal interactions with humans and their environment over time. Formally, a system $S = \{A, H, E\}$ has a quality function $Q_O(t)$ evaluated by every observer $O$; a unilateral interaction changes only one party's future state, while a reciprocal interaction changes both. A meaningful interaction is one whose long-term quality, after a horizon $T_S$, is no worse than if the interaction had not occurred, and a coexisting agent is one whose reciprocal presence is at least as good for the system as its absence. From this the paper derives two properties: situatedness (the agent should improve its own specific system even if the same behavior would harm another) and mutability (the agent and human should mutually shape each other). The paper argues that this definition, if adopted, makes the current stagnant and generic design of embodied agents incompatible with long-term human interaction.
Load-bearing premise
Every definition in the paper rests on there being a quality function $Q_O$ that every observer in the system can evaluate and compare against a counterfactual world in which the agent's interactions never happened; if that comparison cannot be made, whether an agent coexists becomes unverifiable.
Editorial extensions
If this is right
- Long-term in-the-wild deployment of an embodied agent should be treated as a coexistence problem, not a task-completion problem; evaluation should ask whether the human-agent-environment system is better off with the agent than without it.
- Agents will need to be mutable: their behavior, objectives, and even morphology should change through reciprocal interaction with the specific human and environment they live in.
- Situatedness will matter more than generic pretraining: an agent should be evaluated by how much it improves its own system, even if the same behavior would hurt a different household or workplace.
- Current learning paradigms (offline pretraining, meta-learning, continual learning, standard reinforcement learning) are insufficient on their own, because they assume a predefined distribution of novelty; the paper argues for open-ended, human-in-the-loop evolutionary learning instead.
- Foundation models should be used as external stores of generic knowledge to bootstrap situated behavior, not as the agent's internal policy, to avoid freezing generic and stagnant behavior into the system.
Reading between the lines
- The counterfactual in the definition suggests a concrete evaluation protocol the paper does not spell out: measure proxy quality metrics (trust, voluntary interaction frequency, task fluency) with the agent present, then again after a removal period exceeding $T_S$; the definition turns coexistence into an empirical, falsifiable property rather than a metaphor.
- If steamrolling is real, the recursive-data-degradation results cited by the paper imply a technical stability argument for coexistence: a homogeneous fleet of stagnant agents would shrink the distribution of human behavior, degrading future models trained on that behavior, so coexistence would be not only an ethical preference but a way to keep the data ecology healthy.
- The open-system extension in Appendix B suggests coexistence naturally scales beyond a single human and robot to families, teams, and institutions, making the paper's call for ethical and legal frameworks more pressing because the observers who judge quality may not be the same as those who are shaped by the agent.
- A testable extension would be matched long-term deployments of a static optimized agent and a mutable co-shaping agent in similar households, comparing voluntary interaction frequency and user-reported quality after the novelty period; the paper predicts the co-shaping agent retains or grows quality while the static one decays.
Editorial analysis
A structured set of objections, weighed in public.
Referee Report
Summary. This position paper argues that current embodied artificial agents are 'stagnant' and 'generic' because their knowledge is fixed at deployment and drawn from large pre-collected datasets, and that their widespread deployment will therefore 'steamroll' human cultural and behavioral diversity. The authors propose a new property, coexistence, defined as sustained meaningful and reciprocal interactions between an agent, a human user, and the environment, and claim that coexistence is a prerequisite for long-term, in-the-wild human-agent interaction. The paper formalizes this definition in Section 3, draws parallels from developmental biology and research-through-design, and proposes six research directions centered on open-endedness, user-as-designer, morphology, and human-in-the-loop evolution. Appendices discuss scope, measurement proxies, and limitations.
Significance. If the necessity claim were established, the paper would challenge the dominant train-then-deploy paradigm and reframe embodied-agent design around mutual adaptation and situatedness. The manuscript's strengths are its interdisciplinary synthesis of biology, design theory, and long-term HRI, and the concreteness of several proposed directions, especially treating foundation models as external components and involving users in shaping morphology and sensing. The paper is also candid in its appendices about the open problems in measuring its central quantity. However, the central claim is not currently supported: it is either contradicted by the paper's own examples or reduced to a definitional circularity, and the formal apparatus in Section 3 cannot carry the evidential weight placed on it. The paper would be more defensible as a design manifesto or a set of research hypotheses than as an established prerequisite.
major comments (4)
- [Abstract and Section 3 vs. Appendix A] The abstract and Section 3 assert that coexistence is 'a prerequisite for long-term, in-the-wild interaction with humans.' Appendix A, however, cites Sung et al. [130] and Fink et al. [38] as documenting households that use robot vacuum cleaners daily over months, with users rearranging furniture, installing threshold ramps, and changing tidying routines, and then states that these effects 'do not, by themselves, constitute coexistence, but only entanglement.' Under the paper's own criteria, these are long-term, in-the-wild interactions with stagnant, generic agents, which directly contradicts the necessity claim. If the authors reply that these are not 'interactions' in the intended sense, the claim becomes true only by definition, since 'interaction' has been narrowed to mean coexisting interaction; either way, the substantive empirical assertion is lost. The manuscript should either weaken the claim to a design preference or testable hypothesis, or provide an independent characterization of interaction that excludes these cases and explains why that narrower notion is the relevant one.
- [Section 3, Eqs. (3)-(4); Appendices B-C] Equation (3) defines meaningful interaction by comparing Q_O(t' | X_t → Y_t) with Q_O(t' | ∅), where ∅ denotes the absence of interaction. This counterfactual baseline is never defined: there is no specification of how Q_O is evaluated under a counterfactual in which a past interaction did not occur, nor how Q_O is measured and compared across different observers O. Appendix B states that credit assignment is 'an open research problem,' and Appendix C concedes that 'measuring precisely how a user may change as a result of an interaction can be intractable.' Since Eq. (4) inherits these quantities, the formal definition cannot, as it stands, support the claim that coexistence is a prerequisite. In addition, in Eq. (4) the counterfactual term Q_O(t' | H_t ↔ E_t) is evaluated for O = A*, but A* is absent in that counterfactual, so the quantity is undefined for that observer. The authors need to specify an operational or at least well-defined evaluation procedure, or explicitly treat the formalism as illustrative rather than as evidence for the necessity claim.
- [Section 3, Definition of coexistence] There is a definitional circularity in the central claim. Section 3 defines a coexisting agent as one that 'maintain(s) reciprocal and meaningful interactions in the long-term' (Eq. 4), and the abstract then concludes that coexistence is required for long-term, in-the-wild interaction. This makes the conclusion close to analytic: the only interactions counted in the conclusion are those already defined as coexistence. The vacuum-cleaner examples in Appendix A show that the paper needs an independent notion of 'interaction' to make the claim substantive. The authors should either state which observable behaviors count as interaction independently of the Q_O criterion, or explicitly restrict the claim to 'meaningful and reciprocal interaction' and provide an independent justification for why that restricted class is the one that matters for the field's goals.
- [Section 2.3, Steamrolling] Section 2.3 argues that stagnant, generic agents 'will steamroll' human cultures and workflows, and this alleged harm is part of the argument that coexistence is necessary. The supporting evidence is analogical: LLM-style abstracts [44], reduced diversity in LLM-assisted brainstorming [94], and model collapse on recursively generated data [126]. None of these involve embodied agents or long-term human-robot interaction, and the paper acknowledges that direct empirical evidence is limited. Because steamrolling is load-bearing for the normative conclusion, the authors should either present direct empirical evidence from embodied deployments, or explicitly frame steamrolling as an untested risk and derive the recommendation from that weaker, hypothesis-like premise.
minor comments (4)
- [Section 5, direction 4] In the sentence describing the drone study, 'based on its the shape and size' should read 'based on the shape and size.'
- [Appendix D] In the 'Please, just turn on the light' paragraph, 'we are argue that such situated interactions' should read 'we argue that such situated interactions.'
- [References [72] and [73]] References [72] and [73] are identical duplicate entries for Krogh, Markussen, and Bang; one should be removed or the two should be differentiated by content.
- [Acknowledgments] The acknowledgments state 'the Swedish Research Council Swedish Research Council'; the duplicate phrase should be removed.
Circularity Check
The headline 'prerequisite' claim reduces to the paper's own definition; the supporting examples are independent, but the central necessity claim is not derived.
-
self definitional
[Abstract and Section 3 (Definition, Eqs. 1-4); see also Section 7 Conclusion]
"Abstract: 'This paper introduces the concept of coexistence for embodied artificial agents and argues that it is a prerequisite for long-term, in-the-wild interaction with humans.' Section 3: 'An embodied artificial agent is coexisting in a system if it sustains meaningful and reciprocal interactions with humans and their environment over time.' Section 7: 'We proposed coexistence as a new paradigm for the design of embodied agents that emphasizes meaningful, reciprocal interactions sustained over time.'"
Coexistence is defined as sustaining meaningful and reciprocal interactions over time. The central thesis then asserts that coexistence is a prerequisite for long-term, in-the-wild interaction with humans. Read as the same notion, the thesis is true by stipulation: long-term interaction is said to require meaningful, reciprocal, sustained interaction, which is exactly what 'coexistence' was defined to mean. The formal inequality in Eq. 4 restates this condition as QO(t' | A*<->(H,E), H<->E) >= QO(t' | H<->E) without operationalizing QO or the counterfactual baseline; Appendices B and C concede that credit assignment is open and that measuring how a user changes can be intractable.
full rationale
The paper is a position paper, not an empirical derivation, and most of its content is independent of the formal apparatus: the biology and design-theory discussions in Section 4 and the six research directions in Section 5 stand on their own as a design proposal. There is no fitted parameter renamed as a prediction, no imported uniqueness theorem, and the self-citations (e.g., La Delfa et al. [77,78], Leite et al. [84,85]) are used as illustrative examples rather than as load-bearing proofs. The circularity is concentrated in the headline necessity claim: because 'coexistence' is defined as sustaining meaningful and reciprocal interactions over time, the assertion that coexistence is a prerequisite for long-term in-the-wild interaction reduces to the definition if 'interaction' is read narrowly, and it is contradicted by the paper's own vacuum-cleaner examples if read broadly. The formal definition's QO and counterfactual baseline are acknowledged in Appendices B and C to be unoperationalized, so they do not supply independent content. This partial definitional circularity affects the central claim, but the rest of the paper retains independent value, yielding a score of 6.
Assumptions & free parameters
assumptions (6)
- domain assumption In-the-wild interaction is co-constructed with humans, so it cannot be treated as a static optimization problem.
- domain assumption Current embodied agents are stagnant and generic because their knowledge is fixed at training time and derived from large pre-collected datasets.
- domain assumption Steamrolling, the convergence of human culture and workflows toward homogeneous agent-dictated behavior, will occur for embodied agents.
- ad hoc to paper A quality function Q_O can be defined, measured, and compared across observers and under counterfactual absence of interaction.
- domain assumption Analogies from biology (genetic drift, HSP90, digit patterning) and design (double diamond, research through design) transfer to artificial agents.
- domain assumption Human goals and preferences are formed through interaction with the agent rather than known a priori.
Cite this review
Pith. "Pith review of Humans Coexist, So Must Embodied Artificial Agents." pith.science (2026). https://pith.science/paper/NBKRHZNR
@misc{pith2026250204809,
author = {Pith},
title = {Pith review of: Humans Coexist, So Must Embodied Artificial Agents},
year = {2026},
howpublished = {\url{https://pith.science/paper/NBKRHZNR}},
note = {Machine review of arXiv:2502.04809}
}
read the original abstract
This paper introduces the concept of coexistence for embodied artificial agents and argues that it is a prerequisite for long-term, in-the-wild interaction with humans. Contemporary embodied artificial agents excel in static, predefined tasks but fall short in dynamic and long-term interactions with humans. On the other hand, humans can adapt and evolve continuously, exploiting the situated knowledge embedded in their environment and other agents, thus contributing to meaningful interactions. We take an interdisciplinary approach at different levels of organization, drawing from biology and design theory, to understand how human and non-human organisms foster entities that coexist within their specific environments. Finally, we propose key research directions for the artificial intelligence community to develop coexisting embodied agents, focusing on the principles, hardware and learning methods responsible for shaping them.
Figures
Figures from the paper (4 more)
Reference graph
Works this paper leans on
-
[130]
Sung, J., Grinter, R. E., and Christensen, H. I. (2010). Domestic robot ecology. International Journal of Social Robotics, 2(4):417–429. 16
work page 2010
-
[38]
Fink, J., Bauwens, V ., Kaplan, F., and Dillenbourg, P. (2013). Living with a vacuum cleaning robot. International Journal of Social Robotics, 5(3):389–408
2013
-
[44]
Geng, M. and Trotta, R. (2024). Is chatgpt transforming academics’ writing style? arXiv preprint arXiv:2404.08627
arXiv 2024
-
[94]
Meincke, L., Nave, G., and Terwiesch, C. (2025). Chatgpt decreases idea diversity in brain- storming. Nature human behaviour, pages 1–3
2025
-
[126]
Shumailov, I., Shumaylov, Z., Zhao, Y ., Papernot, N., Anderson, R., and Gal, Y . (2024). AI models collapse when trained on recursively generated data. 631(8022):755–759. Publisher: Nature Publishing Group
work page 2024
-
[1]
P., and Singh, S
Abel, D., Barreto, A., Van Roy, B., Precup, D., van Hasselt, H. P., and Singh, S. (2023). A definition of continual reinforcement learning. Advances in Neural Information Processing Systems, 36:50377–50407
2023
-
[2]
L., Almeida, D., Al- tenschmidt, J., Altman, S., Anadkat, S., et al
Achiam, J., Adler, S., Agarwal, S., Ahmad, L., Akkaya, I., Aleman, F. L., Almeida, D., Al- tenschmidt, J., Altman, S., Anadkat, S., et al. (2023). Gpt-4 technical report. arXiv preprint arXiv:2303.08774
arXiv 2023
-
[3]
and Scassellati, B
Admoni, H. and Scassellati, B. (2017). Social eye gaze in human-robot interaction: a review. Journal of Human-Robot Interaction, 6(1):25–63
2017
Show all 160 references
-
[4]
Alonso, E., Jelley, A., Micheli, V ., Kanervisto, A., Storkey, A., Pearce, T., and Fleuret, F. (2024). Diffusion for world modeling: Visual details matter in atari. arXiv preprint arXiv:2405.12399
2024 arXiv
-
[5]
Aly, A., Griffiths, S., and Stramandinoli, F. (2017). Metrics and benchmarks in human-robot interaction: Recent advances in cognitive robotics. Cognitive Systems Research, 43:313–323
2017
-
[6]
Auger, J. (2022). Seven Observations, or Why Domestic Robots are Struggling to Enter the Habitats of Everyday Life. In Meaningful Futures with Robots. Chapman and Hall/CRC
2022
-
[7]
Azeem, R., Hundt, A., Mansouri, M., and Brandão, M. (2024). Llm-driven robots risk enacting discrimination, violence, and unlawful actions
2024
-
[8]
and Schwefel, H.-P
Bäck, T. and Schwefel, H.-P. (1993). An overview of evolutionary algorithms for parameter optimization. Evolutionary computation, 1(1):1–23
1993
-
[9]
Ball, P. (2023). How Life Works: A User’s Guide to the New Biology. Pan Macmillan
2023
-
[10]
Belkaid, M., Kompatsiari, K., De Tommaso, D., Zablith, I., and Wykowska, A. (2021). Mutual gaze with a robot affects human neural activity and delays decision-making processes. Science Robotics, 6(58):eabc5044
2021
-
[11]
and Boer, L
Bewley, H. and Boer, L. (2018). Designing blo-nut: Design principles, choreography and otherness in an expressive social robot. In Proceedings of the 2018 Designing Interactive Sys- tems Conference, DIS ’18, page 1069–1080, New York, NY , USA. Association for Computing Machinery. 10
2018
-
[12]
Bharadhwaj, H. (2024). Position: scaling simulation is neither necessary nor sufficient for in-the-wild robot manipulation. In Forty-first International Conference on Machine Learning
2024
-
[13]
Bianchi, F., Kalluri, P., Durmus, E., Ladhak, F., Cheng, M., Nozza, D., Hashimoto, T., Juraf- sky, D., Zou, J., and Caliskan, A. (2023). Easily accessible text-to-image generation amplifies demographic stereotypes at large scale. In Proceedings of the 2023 ACM Conference on Fa...
2023
-
[14]
Black, K., Brown, N., Driess, D., Esmail, A., Equi, M., Finn, C., Fusai, N., Groom, L., Hausman, K., Ichter, B., et al. (2024). π0: A vision-language-action flow model for general robot control. arXiv preprint arXiv:2410.24164
2024 arXiv
-
[15]
Bongard, J. (2024). The tip and the iceberg: Deep learning and embodiment. Invited talk at CVPR 2024, Summit Flex Hall ABC. Accessed: 2025-05-21
2024
-
[16]
and Elelimy, E
Bowling, M. and Elelimy, E. (2025). Rethinking the foundations for continual reinforcement learning. arXiv preprint arXiv:2504.08161
2025 arXiv
-
[17]
Breazeal, C., Dautenhahn, K., and Kanda, T. (2016). Social robotics. Springer handbook of robotics, pages 1935–1972
2016
-
[18]
G., Gopalakrishnan, K., Han, K., Hausman, K., Herzog, A., Hsu, J., Ichter, B., Irpan, A., Joshi, N., Julian, R., Kalashnikov, D., Kuang, Y ., Leal, I., Lee, L., Lee, T.-W
Brohan, A., Brown, N., Carbajal, J., Chebotar, Y ., Chen, X., Choromanski, K., Ding, T., Driess, D., Dubey, A., Finn, C., Florence, P., Fu, C., Arenas, M. G., Gopalakrishnan, K., Han, K., Hausman, K., Herzog, A., Hsu, J., Ichter, B., Irpan, A., Joshi, N., Julian, R., Kalashnik...
2023 arXiv
-
[19]
Brooks, R. A. (1991). Intelligence without representation. Artificial Intelligence, 47(1-3):139– 159
1991
-
[20]
D., Dhariwal, P., Neelakantan, A., Shyam, P., Sastry, G., Askell, A., et al
Brown, T., Mann, B., Ryder, N., Subbiah, M., Kaplan, J. D., Dhariwal, P., Neelakantan, A., Shyam, P., Sastry, G., Askell, A., et al. (2020). Language models are few-shot learners. Advances in neural information processing systems, 33:1877–1901
2020
-
[21]
D., Edwards, A., Parker-Holder, J., Shi, Y ., Hughes, E., Lai, M., Mavalankar, A., Steigerwald, R., Apps, C., et al
Bruce, J., Dennis, M. D., Edwards, A., Parker-Holder, J., Shi, Y ., Hughes, E., Lai, M., Mavalankar, A., Steigerwald, R., Apps, C., et al. (2024). Genie: Generative interactive en- vironments. In Forty-first International Conference on Machine Learning
2024
-
[22]
Bryden, K. (2006). Using a human-in-the-loop evolutionary algorithm to create data-driven music. In 2006 IEEE International Conference on Evolutionary Computation, pages 2065–2071. IEEE
2006
-
[23]
Chen, M., Nikolaidis, S., Soh, H., Hsu, D., and Srinivasa, S. (2018). Planning with trust for human-robot collaboration. In Proceedings of the 2018 ACM/IEEE international conference on human-robot interaction, pages 307–315
2018
-
[24]
Chen, N., Liu, X., Zhai, Y ., and Hu, X. (2023). Development and validation of a robot social presence measurement dimension scale. Scientific Reports, 13(1):2911
2023
-
[25]
Chen, X., Hu, J., Jin, C., Li, L., and Wang, L. (2021). Understanding domain randomization for sim-to-real transfer. arXiv preprint arXiv:2110.03239
2021 arXiv
-
[26]
Clark, A. (2001). Mindware : an introduction to the philosophy of cognitive science. New York : Oxford University Press
2001
-
[27]
Collins, F. S. and Fink, L. (1995). The human genome project. Alcohol Health and Research World, 19(3):190–195
1995
-
[28]
i choose... you!
Correia, F., Petisca, S., Alves-Oliveira, P., Ribeiro, T., Melo, F. S., and Paiva, A. (2019). “i choose... you!” membership preferences in human–robot teams. Autonomous Robots, 43:359–373
2019
-
[29]
Cross, N. (2000). Engineering design methods: strategies for product design. Wiley, Chichester ; New York, 3rd ed edition
2000
-
[30]
Darlow, L., Regan, C., Risi, S., Seely, J., and Jones, L. (2025). Continuous thought machines. 11
2025
-
[31]
M., Allouch, S
de Graaf, M. M., Allouch, S. B., and van Dijk, J. A. (2016). Long-term acceptance of social robots in domestic environments: insights from a user’s perspective. In 2016 AAAI spring symposium series
2016
-
[32]
Dosovitskiy, A., Beyer, L., Kolesnikov, A., Weissenborn, D., Zhai, X., Unterthiner, T., Dehghani, M., Minderer, M., Heigold, G., Gelly, S., Uszkoreit, J., and Houlsby, N. (2021). An image is worth 16x16 words: Transformers for image recognition at scale
2021
-
[33]
Dourish, P. (2001). Where the Action Is: The Foundations of Embodied Interaction. The MIT Press
2001
-
[34]
S., Lynch, C., Chowdhery, A., Ichter, B., Wahid, A., Tompson, J., Vuong, Q., Yu, T., et al
Driess, D., Xia, F., Sajjadi, M. S., Lynch, C., Chowdhery, A., Ichter, B., Wahid, A., Tompson, J., Vuong, Q., Yu, T., et al. (2023). Palm-e: An embodied multimodal language model. arXiv preprint arXiv:2303.03378
2023 arXiv
-
[35]
and Bailey, T
Durrant-Whyte, H. and Bailey, T. (2006). Simultaneous localization and mapping: part i. IEEE robotics & automation magazine, 13(2):99–110
2006
-
[36]
Dörrenbächer, J., Hassenzahl, M., Neuhaus, R., and Ringfort-Felner, R. (2022). Towards Designing Meaningful Relationships with Robots. In Meaningful Futures with Robots—Designing a New Coexistence, pages 3–29. Chapman and Hall/CRC, Boca Raton, 1 edition
2022
-
[37]
Esterwood, C., Essenmacher, K., Yang, H., Zeng, F., and Robert, L. P. (2021). A meta-analysis of human personality and robot acceptance in human-robot interaction. In Proceedings of the 2021 CHI conference on human factors in computing systems, pages 1–18
2021
-
[39]
Finn, C., Abbeel, P., and Levine, S. (2017). Model-agnostic meta-learning for fast adaptation of deep networks. In International conference on machine learning, pages 1126–1135. PMLR
2017
-
[40]
Firoozi, R., Tucker, J., Tian, S., Majumdar, A., Sun, J., Liu, W., Zhu, Y ., Song, S., Kapoor, A., Hausman, K., et al. (2023). Foundation models in robotics: Applications, challenges, and the future. The International Journal of Robotics Research, page 02783649241281508
2023
-
[41]
and Stankowski, A
Flechtner, R. and Stankowski, A. (2023). Ai is not a wildcard: Challenges for integrating ai into the design curriculum. In Proceedings of the 5th Annual Symposium on HCI Education, EduCHI ’23, page 72–77, New York, NY , USA. Association for Computing Machinery
2023
-
[42]
Frauenberger, C. (2019). Entanglement hci the next wave? ACM Trans. Comput.-Hum. Interact., 27(1)
2019
-
[43]
and Royal College of Art
Frayling, C. and Royal College of Art. (1993). Research in art and design. Number vol. 1, no. 1 in Royal College of Art research papers. Royal College of Art, London. OCLC: 48866129
1993
-
[45]
and Kirschner, M
Gerhart, J. and Kirschner, M. (2007). The theory of facilitated variation. Proceedings of the National Academy of Sciences of the United States of America, 104(Suppl 1):8582–8589
2007
-
[46]
Gillet, S., Vázquez, M., Andrist, S., Leite, I., and Sebo, S. (2024). Interaction-shaping robotics: Robots that influence interactions between other agents. 13(1):12:1–12:23
2024
-
[47]
and Sharot, T
Glickman, M. and Sharot, T. (2024). How human–AI feedback loops alter human perceptual, emotional and social judgements. pages 1–15. Publisher: Nature Publishing Group
2024
-
[48]
spherical human
Gombolay, M. (2024). Human-robot alignment through interactivity and interpretability: Don’t assume a “spherical human”. In Proceedings of the Thirty-Third International Joint Conference on Artificial Intelligence, pages 8523–8528
2024
-
[49]
and Diepold, K
Gronauer, S. and Diepold, K. (2022). Multi-agent deep reinforcement learning: a survey. Artificial Intelligence Review, 55(2):895–943
2022
-
[50]
and Yang, X
Guo, Y . and Yang, X. J. (2021). Modeling and predicting trust dynamics in human–robot teaming: A bayesian inference approach. International Journal of Social Robotics, 13(8):1899– 1909
2021
-
[51]
A., Billings, D
Hancock, P. A., Billings, D. R., Schaefer, K. E., Chen, J. Y ., De Visser, E. J., and Parasuraman, R. (2011). A meta-analysis of factors affecting trust in human-robot interaction. Human factors, 53(5):517–527. 12
2011
-
[52]
Hara, S., Tsuchiya, M., and Kimura, T. (2021). Robust Control of Automatic Low-Speed Driving Motorcycle MOTOROiD. In2021 IEEE 10th Global Conference on Consumer Electronics (GCCE), pages 645–646, Kyoto, Japan. IEEE
2021
-
[53]
N., Contier, O., Teichmann, L., Rockter, A
Hebart, M. N., Contier, O., Teichmann, L., Rockter, A. H., Zheng, C. Y ., Kidder, A., Cor- riveau, A., Vaziri-Pashkam, M., and Baker, C. I. (2023). Things-data, a multimodal collection of large-scale datasets for investigating object representations in human brain and behavior...
2023
-
[54]
Ho, J., Jain, A., and Abbeel, P. (2020). Denoising diffusion probabilistic models. Advances in neural information processing systems, 33:6840–6851
2020
-
[55]
Hoffman, G. (2019). Evaluating fluency in human–robot collaboration. IEEE Transactions on Human-Machine Systems, 49(3):209–218
2019
-
[56]
and Breazeal, C
Hoffman, G. and Breazeal, C. (2007). Effects of anticipatory action on human-robot teamwork efficiency, fluency, and perception of team. In Proceedings of the ACM/IEEE international conference on Human-robot interaction, pages 1–8
2007
-
[57]
Hospedales, T., Antoniou, A., Micaelli, P., and Storkey, A. (2021). Meta-learning in neural networks: A survey. IEEE transactions on pattern analysis and machine intelligence, 44(9):5149– 5169
2021
-
[58]
D., Parker-Holder, J., Behbahani, F., Mavalankar, A., Shi, Y ., Schaul, T., and Rocktäschel, T
Hughes, E., Dennis, M. D., Parker-Holder, J., Behbahani, F., Mavalankar, A., Shi, Y ., Schaul, T., and Rocktäschel, T. (2024). Position: Open-endedness is essential for artificial superhuman intelligence. In Salakhutdinov, R., Kolter, Z., Heller, K., Weller, A., Oliver, N., Sc...
2024
-
[59]
M., White, A., Da Silva, B
Jordan, S. M., White, A., Da Silva, B. C., White, M., and Thomas, P. S. (2024). Position: Benchmarking is limited in reinforcement learning research. arXiv preprint arXiv:2406.16241
2024 arXiv
-
[60]
P., Littman, M
Kaelbling, L. P., Littman, M. L., and Cassandra, A. R. (1998). Planning and acting in partially observable stochastic domains. Artificial intelligence, 101(1-2):99–134
1998
-
[61]
F., and Sabanovi ´c, S
Kamino, W., Jung, M. F., and Sabanovi ´c, S. (2024). Constructing a social life with robots: Shifting away from design patterns towards interaction ritual chains. In Proceedings of the 2024 ACM/IEEE International Conference on Human-Robot Interaction, HRI ’24, page 343–351, Ne...
2024
-
[62]
Kanda, T., Sato, R., Saiwaki, N., and Ishiguro, H. (2007). A two-month field trial in an elemen- tary school for long-term human–robot interaction. IEEE Transactions on robotics, 23(5):962–971
2007
-
[63]
J., Weintrop, D., and Grossman, T
Kazemitabaar, M., Hou, X., Henley, A., Ericson, B. J., Weintrop, D., and Grossman, T. (2024). How novices use llm-based code generators to solve cs1 coding tasks in a self-paced learning environment. In Proceedings of the 23rd Koli Calling International Conference on Comput- i...
2024
-
[64]
Khanna, P., Yadollahi, E., Björkman, M., Leite, I., and Smith, C. (2023). Effects of explanation strategies to resolve failures in human-robot collaboration. In 2023 32nd IEEE International Conference on Robot and Human Interactive Communication (RO-MAN), pages 1829–1836
2023
-
[65]
Khavas, Z. R. (2021). A review on trust in human-robot interaction. arXiv preprint arXiv:2105.10045
2021 arXiv
-
[66]
Kidd, C. D. and Breazeal, C. (2008). Robots at home: Understanding long-term human-robot interaction. In 2008 IEEE/RSJ International Conference on Intelligent Robots and Systems, pages 3230–3235. IEEE
2008
-
[67]
Kim, G., Xiao, C., Konishi, T., and Liu, B. (2023). Learnability and algorithm for continual learning. In International Conference on Machine Learning, pages 16877–16896. PMLR
2023
-
[68]
Kompatsiari, K., Tikhanoff, V ., Ciardo, F., Metta, G., and Wykowska, A. (2017). The importance of mutual gaze in human-robot interaction. In Social Robotics: 9th International Conference, ICSR 2017, Tsukuba, Japan, November 22-24, 2017, Proceedings 9, pages 443–452. Springer
2017
-
[69]
Koskinen, I. K. (2011). Design research through practice: from the lab, field, and showroom. Morgan Kaufmann/Elsevier, Waltham, MA. 13
2011
-
[70]
Kriegman, S. (2020). Design for an Increasingly Protean Machine. Graduate College Disserta- tions and Theses
2020
-
[71]
S., Levin, M., Kramer-Bottiglio, R., and Bongard, J
Kriegman, S., Walker, S., Shah, D. S., Levin, M., Kramer-Bottiglio, R., and Bongard, J. (2019). Automated shapeshifting for function recovery in damaged robots. In Proceedings of Robotics: Science and Systems, FreiburgimBreisgau, Germany
2019
-
[72]
G., Markussen, T., and Bang, A
Krogh, P. G., Markussen, T., and Bang, A. L. (2015a). Ways of drifting—five methods of experimentation in research through design. In Chakrabarti, A., editor, ICoRD’15 – Research into Design Across Boundaries Volume 1, pages 39–50, New Delhi. Springer India
2015
-
[73]
G., Markussen, T., and Bang, A
Krogh, P. G., Markussen, T., and Bang, A. L. (2015b). Ways of drifting—five methods of experimentation in research through design. In Chakrabarti, A., editor, ICoRD’15 – Research into Design Across Boundaries Volume 1, pages 39–50, New Delhi. Springer India
2015
-
[74]
O., Isola, P., and Ha, D
Kumar, A., Lu, C., Kirsch, L., Tang, Y ., Stanley, K. O., Isola, P., and Ha, D. (2024). Automating the search for artificial life with foundation models. arXiv preprint arXiv:2412.17799
2024 arXiv
-
[75]
La Delfa, J. (2023). Cultivating Mechanical Sympathy : Making meaning with ambiguous machines. PhD thesis, KTH, Media Technology and Interaction Design, MID. QC 20230927
2023
-
[76]
A., Luke, E., Koder, B., and Mueller, F
La Delfa, J., Bayta¸ s, M. A., Luke, E., Koder, B., and Mueller, F. F. (2020). Designing drone chi: Unpacking the thinking and making of somaesthetic human-drone interaction. In Proceedings of the 2020 ACM Designing Interactive Systems Conference, DIS ’20, page 575–586, New Yo...
2020
-
[77]
La Delfa, J., Garrett, R., Lampinen, A., and Höök, K. (2024a). How to train your drone: Exploring the umwelt as a design metaphor for human-drone interaction. In Proceedings of the 2024 ACM Designing Interactive Systems Conference, DIS ’24, page 2987–3001, New York, NY , USA. ...
2024
-
[78]
La Delfa, J., Garrett, R., Lampinen, A., and Höök, K. (2024b). Articulating mechanical sympathy for somaesthetic human–machine relations. InProceedings of the 2024 ACM Conference on Designing Interactive Systems, pages 1–18
2024
-
[79]
Laban, G., Kappas, A., Morrison, V ., and Cross, E. S. (2024). Building long-term human–robot relationships: Examining disclosure, perception and well-being across time. International Journal of Social Robotics, 16(5):1–27
2024
-
[80]
Lai, S., Potter, Y ., Kim, J., Zhuang, R., Song, D., and Evans, J. (2024). Position: Evolving AI collectives enhance human diversity and enable self-regulation. In Forty-first International Conference on Machine Learning
2024
-
[81]
and Johnson, M
Lakoff, G. and Johnson, M. (1985). Metaphors we live by. Univ. of Chicago Press, Chicago, Ill., 5. [dr.] edition
1985
-
[82]
J., Sha, F., and Breazeal, C
Lee, J. J., Sha, F., and Breazeal, C. (2019). A bayesian theory of mind approach to nonverbal communication. In 2019 14th ACM/IEEE International Conference on Human-Robot Interaction (HRI), pages 487–496. IEEE
2019
-
[83]
O., and Ziyaee, T
Lehman, J., Meyerson, E., El-Gaaly, T., Stanley, K. O., and Ziyaee, T. (2025). Evolution and the knightian blindspot of machine learning. arXiv preprint arXiv:2501.13075
2025 arXiv
-
[84]
Leite, I., Castellano, G., Pereira, A., Martinho, C., and Paiva, A. (2014). Empathic robots for long-term interaction: evaluating social presence, engagement and perceived support in children. International Journal of Social Robotics, 6:329–341
2014
-
[85]
Leite, I., Martinho, C., and Paiva, A. (2013). Social robots for long-term interaction: a survey. International Journal of Social Robotics, 5:291–308
2013
-
[86]
and Guo, H
Li, K. and Guo, H. (2024). Human-in-the-loop policy optimization for preference-based multi-objective reinforcement learning. arXiv preprint arXiv:2401.02160
2024 arXiv
-
[87]
Li, N., Ma, L., Yu, G., Xue, B., Zhang, M., and Jin, Y . (2023a). Survey on evolutionary deep learning: Principles, algorithms, applications, and open issues. ACM Computing Surveys, 56(2):1–34
2023
-
[88]
X., and Wen, J.-R
Li, Y ., Du, Y ., Zhou, K., Wang, J., Zhao, W. X., and Wen, J.-R. (2023b). Evaluating object hallucination in large vision-language models. InProceedings of the 2023 Conference on Empirical Methods in Natural Language Processing, pages 292–305. 14
2023
-
[89]
A., and Hindriks, K
Ligthart, M., Neerincx, M. A., and Hindriks, K. V . (2019). Getting acquainted for a long-term child-robot interaction. In International Conference on Social Robotics, pages 423–433. Springer
2019
-
[90]
H., Xu, H., and Mannor, S
Lim, S. H., Xu, H., and Mannor, S. (2013). Reinforcement learning in robust markov decision processes. Advances in Neural Information Processing Systems, 26
2013
-
[91]
and Schiphorst, T
Loke, L. and Schiphorst, T. (2018). The somatic turn in human-computer interaction. Interac- tions, 25(5):54–5863
2018
-
[92]
Lu, H., Yang, G., Fei, N., Huo, Y ., Lu, Z., Luo, P., and Ding, M. (2023). Vdt: General-purpose video diffusion transformers via mask modeling. arXiv preprint arXiv:2305.13311
2023 arXiv
-
[93]
McLean, S., Read, G. J. M., Thompson, J., Baber, C., Stanton, N. A., and Salmon, P. M. (2023). The risks associated with artificial general intelligence: A systematic review. 35(5):649–663. Publisher: Taylor & Francis _eprint: https://doi.org/10.1080/0952813X.2021.1964003
2023
-
[95]
S., Heess, N., Guez, A., et al
Mesnard, T., Weber, T., Viola, F., Thakoor, S., Saade, A., Harutyunyan, A., Dabney, W., Stepleton, T. S., Heess, N., Guez, A., et al. (2021). Counterfactual credit assignment in model-free reinforcement learning. In International Conference on Machine Learning, pages 7654–7664. PMLR
2021
-
[96]
Mosqueira-Rey, E., Hernández-Pereira, E., Alonso-Ríos, D., Bobes-Bascarán, J., and Fernández- Leal, Á. (2023). Human-in-the-loop machine learning: a state of the art. Artificial Intelligence Review, 56(4):3005–3054
2023
-
[97]
and Clune, J
Mouret, J. and Clune, J. (2015). Illuminating search spaces by mapping elites. CoRR, abs/1504.04909
2015 arXiv
-
[98]
L., and Prescott, T
Naneva, S., Sarda Gou, M., Webb, T. L., and Prescott, T. J. (2020). A systematic review of attitudes, anxiety, acceptance, and trust towards social robots. International Journal of Social Robotics, 12(6):1179–1201
2020
-
[99]
and Dimitri, N
Naudé, W. and Dimitri, N. (2020). The race for an artificial general intelligence: implications for public policy. 35(2):367–379
2020
-
[100]
Neff, G. (2016). Talking to bots: Symbiotic agency and the case of tay. International journal of Communication
2016
-
[101]
Nikolaidis, S., Hsu, D., and Srinivasa, S. (2017). Human-robot mutual adaptation in col- laborative tasks: Models and experiments. The International Journal of Robotics Research , 36(5-7):618–634
2017
-
[102]
Norman, D. A. (2010). Living with Complexity. MIT Press, Cambridge, MA
2010
-
[103]
and Henriksson, P
Nygren, E. and Henriksson, P. (1992). Reading the medical record. I. Analysis of physicians’ ways of reading the medical record. Computer Methods and Programs in Biomedicine, 39(1-2):1– 12
1992
-
[104]
Odom, W., Wakkary, R., Hol, J., Naus, B., Verburg, P., Amram, T., and Chen, A. Y . S. (2019). Investigating slowness as a frame to design longer-term experiences with personal data: A field study of olly. In Proceedings of the 2019 CHI Conference on Human Factors in Computing ...
2019
-
[105]
Oertel, C., Castellano, G., Chetouani, M., Nasir, J., Obaid, M., Pelachaud, C., and Peters, C. (2020). Engagement in human-agent interaction: An overview. Frontiers in Robotics and AI, 7:92
2020
-
[106]
O’Neill, A., Rehman, A., Gupta, A., Maddukuri, A., Gupta, A., Padalkar, A., Lee, A., Pooley, A., Gupta, A., Mandlekar, A., et al. (2023). Open x-embodiment: Robotic learning datasets and rt-x models. arXiv preprint arXiv:2310.08864
2023 arXiv
-
[107]
Paolo, G., Gonzalez-Billandon, J., and Kégl, B. (2024a). Position: A call for embodied ai. In Forty-first International Conference on Machine Learning
2024
-
[108]
Paolo, G., Gonzalez-Billandon, J., and Kégl, B. (2024b). Position: A call for embodied AI. In Salakhutdinov, R., Kolter, Z., Heller, K., Weller, A., Oliver, N., Scarlett, J., and Berkenkamp, F., editors, Proceedings of the 41st International Conference on Machine Learning, vol...
2024
-
[109]
T., Gillet, S., Winkle, K., and Leite, I
Parreira, M. T., Gillet, S., Winkle, K., and Leite, I. (2023). How did we miss this? a case study on unintended biases in robot social behavior. In Companion of the 2023 ACM/IEEE International Conference on Human-Robot Interaction, HRI ’23, page 11–20, New York, NY , USA. Asso...
2023
-
[110]
and Bongard, J
Pfeifer, R. and Bongard, J. (2006). How the Body Shapes the Way We Think: A New View of Intelligence. A Bradford Book. MIT Press, Cambridge, Massachusetts
2006
-
[111]
Pfeifer, R., Iida, F., and Bongard, J. (2005). New robotics: Design principles for intelligent systems. Artificial Life, 11(1-2):99–120
2005
-
[112]
Pignatelli, E., Ferret, J., Geist, M., Mesnard, T., van Hasselt, H., Pietquin, O., and Toni, L. (2023). A survey of temporal credit assignment in deep reinforcement learning. arXiv preprint arXiv:2312.01072
2023 arXiv
-
[113]
A., Saparov, A., and Mitchell, T
Platanios, E. A., Saparov, A., and Mitchell, T. (2020). Jelly bean world: A testbed for never-ending learning. arXiv preprint arXiv:2002.06306
2020 arXiv
-
[114]
W., Hallacy, C., Ramesh, A., Goh, G., Agarwal, S., Sastry, G., Askell, A., Mishkin, P., Clark, J., et al
Radford, A., Kim, J. W., Hallacy, C., Ramesh, A., Goh, G., Agarwal, S., Sastry, G., Askell, A., Mishkin, P., Clark, J., et al. (2021). Learning transferable visual models from natural language supervision. In International conference on machine learning, pages 8748–8763. PMLR
2021
-
[115]
Rakhymbayeva, N., Amirova, A., and Sandygulova, A. (2021). A long-term engagement with a social robot for autism therapy. Frontiers in Robotics and AI, 8:669972
2021
-
[116]
Raspopovic, J., Marcon, L., Russo, L., and Sharpe, J. (2014). Digit patterning is controlled by a bmp-sox9-wnt turing network modulated by morphogen gradients. Science, 345:566 – 570
2014
-
[117]
Redström, J. (2017). Making design theory. MIT Press, Cambridge, Massachusetts
2017
-
[118]
Reimann, M., van de Graaf, J., van Gulik, N., Van De Sanden, S., Verhagen, T., and Hindriks, K. (2023). Social robots in the wild and the novelty effect. In International Conference on Social Robotics, pages 38–48. Springer
2023
-
[119]
Rodgers, P. (2011). Product design. Portfolio. Laurence King, London
2011
-
[120]
Rutherford, S. L. and Lindquist, S. (1998). Hsp90 as a capacitor for morphological evolution. Nature, 396(6709):336–342
1998
-
[121]
B., Brandstatter, J., Yalcin, U., and Bartneck, C
Sandoval, E. B., Brandstatter, J., Yalcin, U., and Bartneck, C. (2021). Robot likeability and reciprocity in human robot interaction: Using ultimatum game to determinate reciprocal likeable robot strategies. International Journal of Social Robotics, 13(4):851–862
2021
-
[122]
Sandry, E. (2015). Re-evaluating the form and communication of social robots - the benefits of collaborating with machinelike robots. Int. J. Soc. Robotics, 7(3):335–346
2015
-
[123]
Schuhmann, C., Beaumont, R., Vencu, R., Gordon, C., Wightman, R., Cherti, M., Coombes, T., Katta, A., Mullis, C., Wortsman, M., et al. (2022). Laion-5b: An open large-scale dataset for training next generation image-text models. Advances in Neural Information Processing System...
2022
-
[124]
Schön, D. A. (1983). The reflective practitioner: how professionals think in action . Basic Books, New York
1983
-
[125]
Sharp, H., Rogers, Y ., and Preece, J. (2023). Interaction design: beyond human-computer interaction. John Wiley & Sons, Inc, Hoboken, sixth edition edition
2023
-
[127]
J., Guez, A., Sifre, L., Van Den Driessche, G., Schrit- twieser, J., Antonoglou, I., Panneershelvam, V ., Lanctot, M., et al
Silver, D., Huang, A., Maddison, C. J., Guez, A., Sifre, L., Van Den Driessche, G., Schrit- twieser, J., Antonoglou, I., Panneershelvam, V ., Lanctot, M., et al. (2016). Mastering the game of go with deep neural networks and tree search. nature, 529(7587):484–489
2016
-
[128]
Stanley, K. O. and Lehman, J. (2015). Why greatness cannot be planned: the myth of the objective. Springer International Publishing, Cham Heidelberg New York Dordrecht London
2015
-
[129]
Suchman, L. A. (2006). Human-machine reconfigurations plans and situated actions. Cam- bridge University Press, Cambridge ; New York, 2nd ed. edition. OCLC: 1035691746
2006
-
[131]
S., Barto, A
Sutton, R. S., Barto, A. G., et al. (1998). Reinforcement learning: An introduction, volume 1. MIT press Cambridge
1998
-
[132]
D., and Toshev, A
Szot, A., Schwarzer, M., Agrawal, H., Mazoure, B., Metcalf, R., Talbott, W., Mackraz, N., Hjelm, R. D., and Toshev, A. T. (2023). Large language models as generalizable policies for embodied tasks. In The Twelfth International Conference on Learning Representations
2023
-
[133]
Székely, É., Miniot, J., and Hejná, M. (2025). Will ai shape the way we speak? the emerging sociolinguistic influence of synthetic voices
2025
-
[134]
and Lee, J
Tae, M. and Lee, J. (2020). The effect of robot’s ice-breaking humor on likeability and future contact intentions. In Companion of the 2020 ACM/IEEE international conference on human-robot interaction, pages 462–464
2020
-
[135]
S., and Kizilcec, R
Tao, Y ., Viberg, O., Baker, R. S., and Kizilcec, R. F. (2024). Cultural bias and cultural alignment of large language models. PNAS Nexus, 3(9):pgae346
2024
-
[136]
B., Loke, L., Núñez-Pacheco, C., Straker, K., and Wrigley, C
Tomitsch, M., Borthwick, M., Ahmadpour, N., Cooper, C., Frawley, J., Hepburn, L.-A., Kocaballi, A. B., Loke, L., Núñez-Pacheco, C., Straker, K., and Wrigley, C. (2020). Design. Think. Make. Break. Repeat: a handbook of methods. BIS, Amsterdam, revised edition edition
2020
-
[137]
Tonkinwise, C. (2004). Is Design Finished? Dematerialisation and Changing Things. Design Philosophy Papers, 2(3):177–195
2004
-
[138]
A., Najjar, A., and Rodríguez-Lera, F
Tulli, S., Ambrossio, D. A., Najjar, A., and Rodríguez-Lera, F. J. (2019). Great expecta- tions & aborted business initiatives: The paradox of social robot between research and industry. BNAIC/BENELEARN, 1
2019
-
[139]
zero-shot
Udandarao, V ., Prabhu, A., Ghosh, A., Sharma, Y ., Torr, P., Bibi, A., Albanie, S., and Bethge, M. (2024). No" zero-shot" without exponential data: Pretraining concept frequency determines multimodal model performance. In The Thirty-eighth Annual Conference on Neural Informat...
2024
-
[140]
R., and Stone, P
Vasco, M., Seno, T., Kawamoto, K., Subramanian, K., Wurman, P. R., and Stone, P. (2024). A super-human vision-based reinforcement learning agent for autonomous racing in Gran Turismo. Reinforcement Learning Journal, 4:1674–1710
2024
-
[141]
N., Kaiser, L., and Polosukhin, I
Vaswani, A., Shazeer, N., Parmar, N., Uszkoreit, J., Jones, L., Gomez, A. N., Kaiser, L., and Polosukhin, I. (2023). Attention is all you need
2023
-
[142]
Vergunst, J. L. and Ingold, T., editors (2016). Ways of Walking: Ethnography and Practice on Foot. Routledge, London
2016
-
[143]
M., Mathieu, M., Dudzik, A., Chung, J., Choi, D
Vinyals, O., Babuschkin, I., Czarnecki, W. M., Mathieu, M., Dudzik, A., Chung, J., Choi, D. H., Powell, R., Ewalds, T., Georgiev, P., et al. (2019). Grandmaster level in starcraft ii using multi-agent reinforcement learning. nature, 575(7782):350–354
2019
-
[144]
Wang, L., Zhang, X., Su, H., and Zhu, J. (2024). A comprehensive survey of continual learning: Theory, method and application. IEEE Transactions on Pattern Analysis and Machine Intelligence
2024
-
[145]
Wang, R., Zhang, J., Chen, J., Xu, Y ., Li, P., Liu, T., and Wang, H. (2023a). Dexgraspnet: A large-scale robotic dexterous grasp dataset for general objects based on simulation. In 2023 IEEE International Conference on Robotics and Automation (ICRA), pages 11359–11366. IEEE
2023
-
[146]
Wang, Z., Du, Y ., Zhang, Y ., Fang, M., and Huang, B. (2023b). Macca: Offline multi-agent reinforcement learning with causal credit assignment. arXiv preprint arXiv:2312.03644
2023 arXiv
-
[147]
and Flores, F
Winograd, T. and Flores, F. (1986). Understanding Computers and Cognition: A New Foundation for Design. Language and being. Ablex Publishing Corporation
1986
-
[148]
Xiang, J., Tao, T., Gu, Y ., Shu, T., Wang, Z., Yang, Z., and Hu, Z. (2024). Language models meet world models: Embodied experiences enhance language models. Advances in neural information processing systems, 36
2024
-
[149]
Yahara, I. (1999). The role of HSP90 in evolution. Genes to Cells, 4(7):375–379
1999
-
[150]
Yan, D. (2023). Impact of chatgpt on learners in a l2 writing practicum: An exploratory investigation. Education and Information Technologies, 28(11):13943–13967
2023
-
[151]
S., Vu, B., Sharma, A., Bohg, J., and Finn, C
Yang, J., Mark, M. S., Vu, B., Sharma, A., Bohg, J., and Finn, C. (2024a). Robot fine-tuning made easy: Pre-training rewards and policies for autonomous real-world reinforcement learning. In 2024 IEEE International Conference on Robotics and Automation (ICRA), pages 4804–4811. 17
2024
-
[152]
Yang, Z., Liu, A., Liu, Z., Liu, K., Xiong, F., Wang, Y ., Yang, Z., Hu, Q., Chen, X., Zhang, Z., Luo, F., Guo, Z., Li, P., and Liu, Y . (2024b). Position: Towards unified alignment between agents, humans, and environment. In Proceedings of the 41st International Conference on...
2024
-
[153]
Yarats, D., Kostrikov, I., and Fergus, R. (2021). Image augmentation is all you need: Reg- ularizing deep reinforcement learning from pixels. In International conference on learning representations
2021
-
[154]
Yuan, Y ., Tang, K., Shen, J., Zhang, M., and Wang, C. (2024). Measuring social norms of large language models. arXiv preprint arXiv:2404.02491
2024 arXiv
-
[155]
Zanzotto, F. M. (2019). Human-in-the-loop artificial intelligence. Journal of Artificial Intelligence Research, 64:243–252
2019
-
[156]
Zeng, K.-H., Zhang, Z., Ehsani, K., Hendrix, R., Salvador, J., Herrasti, A., Girshick, R., Kembhavi, A., and Weihs, L. (2024). Poliformer: Scaling on-policy rl with transformers results in masterful navigators. arXiv preprint arXiv:2406.20083
2024 arXiv
-
[157]
Zhai, X., Wang, X., Mustafa, B., Steiner, A., Keysers, D., Kolesnikov, A., and Beyer, L. (2022). Lit: Zero-shot transfer with locked-image text tuning. In Proceedings of the IEEE/CVF conference on computer vision and pattern recognition, pages 18123–18133
2022
-
[158]
Zhang, Y ., Beskow, J., and Kjellström, H. (2017). Look but don’t stare: Mutual gaze interaction in social robots. In Social Robotics: 9th International Conference, ICSR 2017, Tsukuba, Japan, November 22-24, 2017, Proceedings 9, pages 556–566. Springer
2017
-
[159]
Zhang, Y ., Li, Y ., Cui, L., Cai, D., Liu, L., Fu, T., Huang, X., Zhao, E., Zhang, Y ., Chen, Y ., et al. (2023). Siren’s song in the ai ocean: a survey on hallucination in large language models. arXiv preprint arXiv:2309.01219
2023 arXiv
-
[160]
turn taking
Zimmerman, J., Forlizzi, J., and Evenson, S. (2007). Research through design as a method for interaction design research in hci. In Proceedings of the SIGCHI Conference on Human Factors in Computing Systems, CHI ’07, page 493–502, New York, NY , USA. Association for Computing ...
2007
Reviewed August 8, 2026 · model on record in the stance chip above.
Discussion (0). Continue with ORCID to comment.