К книге
Эти странные новые разумы: Как ИИ научился говорить и что это значитСписок литературы.
94%
Список литературы.
49

Aher, G., Arriaga, R. I., and Kalai, A. T. (2023), ‘Using Large Language Models to Simulate Multiple Humans and Replicate Human Subject Studies’. arXiv. Available at http://arxiv.org/abs/2208.10264 (accessed 19 October 2023).

Anderson, P. W. (1972), ‘More is Different: Broken Symmetry and the Nature of the Hierarchical Structure of Science’, Science, 177(4047), pp. 393–6. Available at https://doi.org/10.1126/science.177.4047.393.

Aral, S. (2020), The Hype Machine. New York: Currency.

Arcera y Arcas, B. (2022), ‘Do Large Language Models Understand Us?’, Daedalus, 151(2), pp. 183–97. Available at https://doi.org/10.1162/daed_a_01909.

Argyle, L. P. et al. (2023), ‘Out of One, Many: Using Language Models to Simulate Human Samples’, Political Analysis, 31(3), pp. 337–51. Available at https://doi.org/10.1017/pan.2023.2.

Bai, H. et al. (2023), ‘Artificial Intelligence Can Persuade Humans on Political Issues’. Preprint. Open Science Framework. Available at https://doi.org/10.31219/osf.io/stakv.

Bai, Y. et al. (2022), ‘Constitutional AI: Harmlessness from AI Feedback’. arXiv. Available at http://arxiv.org/abs/2212.08073 (accessed 25 October 2023).

Baria, A. T. and Cross, K. (2021), ‘The Brain is a Computer is a Brain: Neuroscience’s Internal Debate and the Social Significance of the Computational Metaphor’. Available at https://doi.org/10.48550/arXiv.2107.14042.

Baroni, M. (2021), ‘On the Proper Role of Linguistically Oriented Deep Net Analysis in Linguistic Theorizing’. Available at https://doi.org/10.48550/arXiv.2106.08694.

Belkin, M. et al. (2019), ‘Reconciling Modern Machine-Learning Practice and the Classical Bias–Variance Trade-Off’, Proceedings of the National Academy of Sciences, 116(32), pp. 15849–54. Available at https://doi.org/10.1073/pnas.1903070116.

Bender, E. M. et al. (2021), ‘On the Dangers of Stochastic Parrots: Can Language Models be Too Big? ’, Proceedings of the 2021 ACM Conference on Fairness, Accountability, and Transparency. FAccT ’21: 2021 ACM Conference on Fairness, Accountability, and Transparency, Virtual Event Canada: ACM, pp. 610–23. Available at https://doi.org/10.1145/3442188.3445922.

Bender, E. M. and Koller, A. (2020), ‘Climbing Towards NLU: On Meaning, Form, and Understanding in the Age of Data’, Proceedings of the 58th Annual Meeting of the Association for Computational Linguistics, pp. 5185–98. Available at https://doi.org/10.18653/v1/2020.acl-main.463.

Bengio, Yoshua, Ducharme, Réjean, Vincent, Pascal, and Jauvin, Christian (2003), ‘A Neural Probabilistic Language Model’, Journal of Machine Learning Research 3, pp. 1137–55, www.jmlr.org/papers/volume3/bengio03a/bengio03a.pdf.

Binz, M. et al. (2023), ‘Meta-Learned Models of Cognition’. arXiv. Available at http://arxiv.org/abs/2304.06729 (accessed 30 October 2023).

Bleses, D., Basbøll, H., and Vach, W. (2011), ‘Is Danish Difficult to Acquire? Evidence from Nordic Past-Tense Studies’, Language and Cognitive Processes, 26(8), pp. 1193–231. Available at https://doi.org/10.1080/01690965.2010.515107.

Bostrom, N. (2014), Superintelligence: Paths, Dangers, Strategies. Oxford: Oxford University Press.

Bottou, L. and Schölkopf, B. (2023), ‘Borges and AI’. arXiv. Available at http://arxiv.org/abs/2310.01425 (accessed 6 October 2023).

Bubeck, S. et al. (2023), ‘Sparks of Artificial General Intelligence: Early Experiments with GPT-4’. arXiv. Available at http://arxiv.org/abs/2303.12712 (accessed 18 February 2024).

Cerina, R. and Duch, R. (2023), ‘Artificially Intelligent Opinion Polling’. arXiv. Available at http://arxiv.org/abs/2309.06029 (accessed 20 October 2023).

Chater, N. and Christiansen, M. (2022), The Language Game: How Improvisation Created Language and Changed the World. London: Bantam Press.

Chen, C. and Shu, K. (2023), ‘Can LLM-Generated Misinformation be Detected?’ arXiv. Available at http://arxiv.org/abs/2309.13788 (accessed 6 October 2023).

Chomsky, Noam (1957), Syntactic Structures, The Hague: Mouton.

Christian, B. (2020), The Alignment Problem: Machine Learning and Human Values. New York: W. W. Norton & Co.

Cobb, M. (2021), The Idea of the Brain: A History. London: Profile.

Cristia, A. et al. (2019), ‘Child-Directed Speech is Infrequent in a Forager-Farmer Population: A Time Allocation Study’, Child Development, 90(3), pp. 759–73. Available at https://doi.org/10.1111/cdev.12974.

Dasgupta, I. et al. (2023), ‘Collaborating with Language Models for Embodied Reasoning’. arXiv. Available at http://arxiv.org/abs/2302.00763 (accessed 17 December 2023).

Davis, M. (2000), The Universal Computer: The Road from Leibniz to Turing. New York: W. W. Norton & Co.

Dawkins, R. (2016), The Blind Watchmaker: Why the Evidence of Evolution Reveals a Universe Without Design. London: Penguin.

De Graaf, M. M. A., Hindriks, F. A., and Hindriks, K. V. (2022), ‘Who Wants to Grant Robots Rights?’, Frontiers in Robotics and AI, 8, 781985. Available at https://doi.org/10.3389/frobt.2021.781985.

Dehaene, S. et al. (2022), ‘Symbols and Mental Programs: A Hypothesis About Human Singularity’, Trends in Cognitive Sciences, 26(9), pp. 751–66. Available at https://doi.org/10.1016/j.tics.2022.06.010.

DeLeo, M. and Guven, E. (2022), ‘Learning Chess with Language Models and Transformers’. arXiv. Available at https://doi.org/10.48550/arXiv.2209.11902.

Depounti, I., Saukko, P., and Natale, S. (2023), ‘Ideal Technologies, Ideal Women: AI and Gender Imaginaries in Redditors’ Discussions on the Replika Bot Girlfriend’, Media, Culture & Society, 45(4), pp. 720–36. Available at https://doi.org/10.1177/01634437221119021.

Downing, T. (2018), 1983: The World at the Brink. London: Little, Brown.

Elkins, K. and Chun, J. (2020), ‘Can GPT-3 Pass a Writer’s Turing Test?’, Journal of Cultural Analytics, 5(2). Available at https://doi.org/10.22148/001c.17212.

Ernst, G. W. and Newell, A. (1967), ‘Some Issues of Representation in a General Problem Solver’, Proceedings of the April 18–20, 1967, Spring Joint Computer Conference on - AFIPS ’67, Atlantic City: ACM Press, pp. 583–600. Available at https://doi.org/10.1145/1465482.1465579.

Feng, X. et al. (2023), ‘ChessGPT: Bridging Policy Learning and Language Modeling’. arXiv. Available at http://arxiv.org/abs/2306.09200 (accessed 1 December 2023).

Gao, L. et al. (2023), ‘PAL: Program-Aided Language Models’. arXiv. Available at http://arxiv.org/abs/2211.10435 (accessed 13 December 2023).

Gardner, R. A. and Gardner, B. T. (1969), ‘Teaching Sign Language to a Chimpanzee: A Standardized System of Gestures Provides a Means of Two-Way Communication with a Chimpanzee’, Science, 165, pp. 664–72. Available at https://doi.org/10.1126/science.165.3894.664.

Gehman, S. et al. (2020), ‘RealToxicityPrompts: Evaluating Neural Toxic Degeneration in Language Models’, in Findings of the Association for Computational Linguistics: EMNLP 2020, pp. 3356–69. Available at https://doi.org/10.18653/v1/2020.findings-emnlp.301.

Glaese, A. et al. (2022), ‘Improving Alignment of Dialogue Agents via Targeted Human Judgements’. arXiv. Available at http://arxiv.org/abs/2209.14375 (accessed 22 October 2023).

Gunkel, D. J. (2018), Robot Rights. Cambridge, MA: MIT Press.

Hackenburg, K. et al. (2023), ‘Comparing the Persuasiveness of Role-Playing Large Language Models and Human Experts on Polarized U.S. Political Issues’. Preprint. Open Science Framework. Available at https://doi.org/10.31219/osf.io/ey8db.

Harari, Y. N. (2015), Sapiens: A Brief History of Humankind. London: Vintage.

Harris, R. A. (2021), The Linguistics Wars. New York: Oxford University Press.

Hartmann, J., Schwenzow, J., and Witte, M. (2023), ‘The Political Ideology of Conversational AI: Converging Evidence on ChatGPT’s Pro-Environmental, Left-Libertarian Orientation’. arXiv. Available at http://arxiv.org/abs/2301.01768 (accessed 20 October 2023).

Hasher, L., Goldstein, D., and Toppino, T. (1977), ‘Frequency and the Conference of Referential Validity’, Journal of Verbal Learning and Verbal Behavior, 16(1), pp. 107–12. Available at https://doi.org/10.1016/S0022-5371(77)80012-1.

Hendrycks, D. (2023), ‘Natural Selection Favors AIs over Humans’. arXiv. Available at http://arxiv.org/abs/2303.16200 (accessed 16 December 2023).

Hendrycks, D., Mazeika, M., and Woodside, T. (2023), ‘An Overview of Catastrophic AI Risks’. arXiv. Available at http://arxiv.org/abs/2306.12001 (accessed 17 December 2023).

Hintzman, D. L. and Ludlam, G. (1980), ‘Differential Forgetting of Prototypes and Old Instances: Simulation by an Exemplar-Based Classification Model’, Memory & Cognition, 8(4), pp. 378–82. Available at https://doi.org/10.3758/BF03198278.

Hochreiter, S. and Schmidhuber, J. (1997), ‘Long Short-Term Memory’, Neural Computation, 9(8), pp. 1735–80. Available at https://doi.org/10.1162/neco.1997.9.8.1735.

Jiang, G. et al. (2023), ‘Evaluating and Inducing Personality in Pre-trained Language Models’. arXiv. Available at http://arxiv.org/abs/2206.07550 (accessed 24 October 2023).

Johnson, M. et al. (2017), ‘Google’s Multilingual Neural Machine Translation System: Enabling Zero-Shot Translation’. arXiv. Available at http://arxiv.org/abs/1611.04558 (accessed 18 May 2023).

Kahneman, D. (2012), Thinking, Fast and Slow. London: Penguin.

Karinshak, E. et al. (2023), ‘Working with AI to Persuade: Examining a Large Language Model’s Ability to Generate Pro-Vaccination Messages’, Proceedings of the ACM on Human-Computer Interaction, 7(CSCW1), pp. 1–29. Available at https://doi.org/10.1145/3579592.

Kim, G., Baldi, P., and McAleer, S. (2023), ‘Language Models Can Solve Computer Tasks’. arXiv. Available at http://arxiv.org/abs/2303.17491 (accessed 11 December 2023).

Klessinger, N., Szczerbinski, M., and Varley, R. (2007), ‘Algebra in a Man with Severe Aphasia’, Neuropsychologia, 45(8), pp. 1642–8. Available at https://doi.org/10.1016/j.neuropsychologia.2007.01.005.

Kocijan, V. et al. (2023), ‘The Defeat of the Winograd Schema Challenge’. arXiv. Available at http://arxiv.org/abs/2201.02387 (accessed 17 February 2024).

Kojima, T. et al. (2023), ‘Large Language Models Are Zero-Shot Reasoners’. arXiv. Available at http://arxiv.org/abs/2205.11916 (accessed 11 December 2023).

Krueger, D., Maharaj, T., and Leike, J. (2020), ‘Hidden Incentives for Auto-Induced Distributional Shift’. arXiv. Available at http://arxiv.org/abs/2009.09153 (accessed 24 November 2023).

Lenat, D. (2022), ‘Creating a 30-Million-Rule System: MCC and Cycorp’, IEEE Annals of the History of Computing, 44(1), pp. 44–56. Available at https://doi.org/10.1109/MAHC.2022.3149468.

Lewis, P. et al. (2021), ‘Retrieval-Augmented Generation for Knowledge-Intensive NLP Tasks’. arXiv. Available at http://arxiv.org/abs/2005.11401 (accessed 9 December 2023).

Li, Y. et al. (2022), ‘Competition-Level Code Generation With AlphaCode’, Science, 378(6624), pp. 1092–7. Available at https://doi.org/10.1126/science.abq1158.

Lin, S., Hilton, J., and Evans, O. (2022), ‘TruthfulQA: Measuring How Models Mimic Human Falsehoods’. arXiv. Available at http://arxiv.org/abs/2109.07958 (accessed 7 October 2023).

Linzen, T., Dupoux, E., and Goldberg, Y. (2016), ‘Assessing the Ability of LSTMs to Learn Syntax-Sensitive Dependencies’, Transactions of the Association for Computational Linguistics, 4, pp. 521–35. Available at https://doi.org/10.1162/tacl_a_00115.

Liu, T. and Low, B. K. H. (2023), ‘Goat: Fine-Tuned LLaMA Outperforms GPT-4 on Arithmetic Tasks’. Available at https://doi.org/10.48550/arXiv.2305.14201.

Lobina, D. (2023), ‘Artificial Intelligence [sic: Machine Learning] and the Best Game in Town; or How Some Philosophers, and the Bbs, Missed a Step’, 3 Quarks Daily, 13 February. Available at https://3quarksdaily.com/3quarksdaily/2023/02/artificial-intelligence-sic-machine-learning-and-the-best-game-in-town-or-how-some-philosophers-and-the-bbs-missed-a-step.html.

Lu, Y., Yu, J., and Huang, S.-H. S. (2023), ‘Illuminating the Black Box: A Psychometric Investigation into the Multifaceted Nature of Large Language Models’. arXiv. Available at http://arxiv.org/abs/2312.14202 (accessed 17 February 2024).

Luccioni, A. S. and Viviano, J. D. (2021), ‘What’s in the Box? A Preliminary Analysis of Undesirable Content in the Common Crawl Corpus’. arXiv. Available at http://arxiv.org/abs/2105.02732 (accessed 6 October 2023).

Luria, A. R., Tsvetkova, L. S., and Futer, D. S. (1965), ‘Aphasia in a Composer’, Journal of the Neurological Sciences, 2(3), pp. 288–92. Available at https://doi.org/10.1016/0022-510X(65)90113-9.

Madaan, A. et al. (2022), ‘Language Models of Code Are Few-Shot Commonsense Learners’. arXiv. Available at https://doi.org/10.48550/arXiv.2210.07128.

Mahowald, K. et al. (2023), ‘Dissociating Language and Thought in Large Language Models: A Cognitive Perspective’. arXiv. Available at http://arxiv.org/abs/2301.06627 (accessed 16 September 2023).

Marcus, G. (2020), ‘The Next Decade in AI: Four Steps Towards Robust Artificial Intelligence’. Preprint. arXiv. Available at http://arxiv.org/abs/2002.06177 (accessed 8 April 2021).

Matz, S. et al. (2023), ‘The Potential of Generative AI for Personalized Persuasion at Scale’. Preprint. PsyArXiv. Available at https://doi.org/10.31234/osf.io/rn97c.

McCulloch, W. S. and Pitts, W. (1943), ‘A Logical Calculus of the Ideas Immanent in Nervous Activity’, Bulletin of Mathematical Biophysics, 5, pp. 115–33. Available at https://doi.org/10.1007/BF02478259.

Metzinger, T. (2021), ‘Artificial Suffering: An Argument for a Global Moratorium on Synthetic Phenomenology’, Journal of Artificial Intelligence and Consciousness, 08(01), pp. 43–66. Available at https://doi.org/10.1142/S270507852150003X.

Mialon, G. et al. (2023), ‘Augmented Language Models: A Survey’. arXiv. Available at http://arxiv.org/abs/2302.07842 (accessed 11 December 2023).

Michel, J.-B. et al. (2011), ‘Quantitative Analysis of Culture Using Millions of Digitized Books’, Science, 331(6014), pp. 176–82. Available at https://doi.org/10.1126/science.1199644.

Mikolov, T. et al. (2013), ‘Distributed Representations of Words and Phrases and Their Compositionality’. arXiv. Available at http://arxiv.org/abs/1310.4546 (accessed 18 February 2024).

Miller, B. A. P. (2015), ‘Automatic Detection of Comment Propaganda in Chinese Media’. Preprint. SSRN Electronic Journal. Available at https://doi.org/10.2139/ssrn.2738325.

Minsky, M. and Papert, S. (1969), Perceptrons: An Introduction to Computational Geometry. Cambridge, Ma: MIT Press.

Moskal, S. et al. (2023), ‘LLMs Killed the Script Kiddie: How Agents Supported by Large Language Models Change the Landscape of Network Threat Testing’. arXiv. Available at http://arxiv.org/abs/2310.06936 (accessed 17 December 2023).

Mosteller, F. and Wallace, D. L. (1963), ‘Inference in an Authorship Problem’, Journal of the American Statistical Association, 58(302), p. 275. Available at https://doi.org/10.2307/2283270.

Nakano, R. et al. (2022), ‘WebGPT: Browser-Assisted Question-Answering with Human Feedback’. arXiv. Available at http://arxiv.org/abs/2112.09332 (accessed 9 December 2023).

Newell, A., Shaw, J. C., and Simon, H. A. (1959), ‘Report on a General Problem-Solving Program’. Available at http://bitsavers.informatik.uni-stuttgart.de/pdf/rand/ipl/P-1584_Report_On_A_General_Problem-Solving_Program_Feb59.pdf.

Noy, S. and Zhang, W. (2023), ‘Experimental Evidence on the Productivity Effects of Generative Artificial Intelligence’, Science, 381(6654), pp. 187–92. Available at https://doi.org/10.1126/science.adh2586.

OpenAI (2023), ‘GPT-4 Technical Report’. arXiv. Available at http://arxiv.org/abs/2303.08774 (accessed 7 October 2023).

Ord, Toby (2020), The Precipice: Existential Risk and the Future of Humanity. London: Bloomsbury.

Ouyang, L. et al. (2022), ‘Training Language Models to Follow Instructions with Human Feedback’. arXiv. Available at http://arxiv.org/abs/2203.02155 (accessed 26 November 2022).

Owen, C. M., Howard, A., and Binder, D. K. (2009), ‘Hippocampus Minor, Calcar Avis, and the Huxley–Owen Debate’, Neurosurgery, 65(6), pp. 1098–105. Available at https://doi.org/10.1227/01.NEU.0000359535.84445.0B.

Pan, Y. et al. (2023), ‘On the Risk of Misinformation Pollution with Large Language Models’. arXiv. Available at http://arxiv.org/abs/2305.13661 (accessed 26 October 2023).

Pariser, Eli (2011), The Filter Bubble: What the Internet is Hiding from You. London: Viking.

Patterson, F. G. (1978), ‘The Gestures of a Gorilla: Language Acquisition in Another Pongid’, Brain and Language, 5(1), pp. 72–97. Available at https://doi.org/10.1016/0093-934X(78)90008-1.

Perez, E. et al. (2022), ‘Discovering Language Model Behaviors with Model-Written Evaluations’. arXiv. Available at http://arxiv.org/abs/2212.09251 (accessed 22 October 2023).

Phuong, M. and Hutter, M. (2022), ‘Formal Algorithms for Transformers’. Available at https://doi.org/10.48550/arXiv.2207.09238.

Piantadosi, S. T. (2023), ‘Modern Language Models Refute Chomsky’s Approach to Language’. LingBuzz. Available at https://lingbuzz.net/lingbuzz/007180.

Piantadosi, S. T. and Hill, F. (2022), ‘Meaning Without Reference in Large Language Models’. Available at https://doi.org/10.48550/arXiv.2208.02957.

Press, O. et al. (2023), ‘Measuring and Narrowing the Compositionality Gap in Language Models’. arXiv. Available at http://arxiv.org/abs/2210.03350 (accessed 13 December 2023).

Ravuri, S. et al. (2021), ‘Skilful Precipitation Nowcasting Using Deep Generative Models of Radar’, Nature, 597(7878), pp. 672–7. Available at https://doi.org/10.1038/s41586-021-03854-z.

Runciman, D. (2019), How Democracy Ends. London: Profile.

Russell, S. (2019), Human Compatible: AI and the Problem of Control. New York: Viking.

Russell, S. and Norvig, P. (2020), Artificial Intelligence: A Modern Approach, 4th edn. Hoboken, Nj: Pearson.

Ryle, G. (2009), The Concept of Mind. London: Routledge.

Sahlgren, M. and Carlsson, F. (2021), ‘The Singleton Fallacy: Why Current Critiques of Language Models Miss the Point’, Frontiers in Artificial Intelligence, 4, 682578. Available at https://doi.org/10.3389/frai.2021.682578.

Santurkar, S. et al. (2023), ‘Whose Opinions Do Language Models Reflect?’ arXiv. Available at http://arxiv.org/abs/2303.17548 (accessed 19 October 2023).

Scheurer, J. et al. (2022), ‘Training Language Models with Language Feedback’. arXiv. Available at http://arxiv.org/abs/2204.14146 (accessed 13 December 2023).

Schick, T. et al. (2023), ‘Toolformer: Language Models Can Teach Themselves to Use Tools’. arXiv. Available at http://arxiv.org/abs/2302.04761 (accessed 17 March 2024).

Searle, J. (1999), ‘The Chinese Room’, in R. A. Wilson and F. C. Keil (eds.), The MIT Encyclopedia of the Cognitive Sciences, Cambridge, Ma: MIT Press.

Sejnowski, T. J. (2020), ‘The Unreasonable Effectiveness of Deep Learning in Artificial Intelligence’, Proceedings of the National Academy of Sciences, 117(48), pp. 30033–8. Available at https://doi.org/10.1073/pnas.1907373117.

Shah, C. and Bender, E. M. (2022), ‘Situating Search’, in ACM SIGIR Conference on Human Information Interaction and Retrieval. CHIIR ’22, Regensburg: ACM, pp. 221–32. Available at https://doi.org/10.1145/3498366.3505816.

Shanahan, M. (2022), ‘Talking About Large Language Models’. arXiv. Available at https://doi.org/10.48550/arXiv.2212.03551.

Shiffrin, R. M. and Schneider, W. (1977), ‘Controlled and Automatic Human Information Processing: II. Perceptual Learning, Automatic Attending and a General Theory’, Psychological Review, 84(2), pp. 127–90. Available at https://doi.org/10.1037/0033-295X.84.2.127.

Skjuve, M. et al. (2021), ‘My Chatbot Companion – A Study of Human–Chatbot Relationships’, International Journal of Human-Computer Studies, 149, 102601. Available at https://doi.org/10.1016/j.ijhcs.2021.102601.

Spatz, H. Ch., Emanns, A., and Reichert, H. (1974), ‘Associative Learning of Drosophila melanogaster’, Nature, 248(5446), pp. 359–61. Available at https://doi.org/10.1038/248359a0.

Sunstein, C. R. (2021), ‘Manipulation as Theft’. Preprint. SSRN Electronic Journal. Available at https://doi.org/10.2139/ssrn.3880048.

Sutskever, I., Vinyals, O., and Le, Q. V. (2014), ‘Sequence to Sequence Learning with Neural Networks’. arXiv. Available at https://doi.org/10.48550/arXiv.1409.3215.

Sutton, R. S. and Barto, A. G. (1998), Reinforcement Learning: An Introduction. Cambridge, MA: MIT Press.

Talland, G. A. (1961), ‘Confabulation in the Wernicke-Korsakoff Syndrome’, Journal of Nervous and Mental Disease, 132(5), pp. 361–81. Available at https://doi.org/10.1097/00005053-196105000-00001.

Tegmark, M. (2017), Life 3.0: Being Human in the Age of Artificial Intelligence. London: Penguin.

Terrace, H. et al. (1979), ‘Can an Ape Create a Sentence?’, Science, 206(4421), pp. 891–902. Available at https://doi.org/10.1126/science.504995.

Thoppilan, R. et al. (2022), ‘LaMDA: Language Models for Dialog Applications’. arXiv. Available at https://doi.org/10.48550/arXiv.2201.08239.

Touvron, H. et al. (2023), ‘LLaMA: Open and Efficient Foundation Language Models’. arXiv. Available at http://arxiv.org/abs/2302.13971 (accessed 23 October 2023).

Turner, M. S. et al. (2008), ‘Confabulation: Damage to a Specific Inferior Medial Prefrontal System’, Cortex, 44(6), pp. 637–48. Available at https://doi.org/10.1016/j.cortex.2007.01.002.

Ullman, T. (2023), ‘Large Language Models Fail on Trivial Alterations to Theory-of-Mind Tasks’, arXiv. Available at https://arxiv.org/pdf/2302.08399.

Vaswani, A. et al. (2017), ‘Attention Is All You Need’. Preprint. arXiv. Available at http://arxiv.org/abs/1706.03762 (accessed 30 October 2020).

Verzijden, M. N. et al. (2015), ‘Male Drosophila melanogaster Learn to Prefer an Arbitrary Trait Associated with Female Mating Status’, Current Zoology, 61(6), pp. 1036–42. Available at https://doi.org/10.1093/czoolo/61.6.1036.

Wallace, E. et al. (2022), ‘Automated Crossword Solving’. arXiv. Available at https://doi.org/10.48550/arXiv.2205.09665.

Wang, W. Y. (2017), ‘ “Liar, Liar, Pants on Fire”: A New Benchmark Dataset for Fake News Detection’, Proceedings of the 55th Annual Meeting of the Association for Computational Linguistics, vol. 2: Short Papers, Vancouver: Association for Computational Linguistics, pp. 422–6. Available at https://doi.org/10.18653/v1/P17-2067.

Webb, T. et al. (2023), ‘A Prefrontal Cortex-Inspired Architecture for Planning in Large Language Models’. arXiv. Available at http://arxiv.org/abs/2310.00194 (accessed 10 December 2023).

Weizenbaum, J. (1966), ‘ELIZA – A Computer Program for the Study of Natural Language Communication Between Man and Machine’, Communications of the ACM, 9(1), pp. 36–45. Available at https://doi.org/10.1145/365153.365168.

Winding, M. et al. (2023), ‘The Connectome of an Insect Brain’, Science, 379(6636). Available at https://doi.org/10.1126/science.add9330.

Winograd, T. (1972), ‘Understanding Natural Language’, Cognitive Psychology, 3(1), pp. 1–191. Available at https://doi.org/10.1016/0010-0285(72)90002-3.

Yang, C. et al. (2023), ‘Large Language Models as Optimizers’. arXiv. Available at http://arxiv.org/abs/2309.03409 (accessed 21 December 2023).

Yang, Z. et al. (2018), ‘HotpotQA: A Dataset for Diverse, Explainable Multi-Hop Question Answering’. arXiv. Available at http://arxiv.org/abs/1809.09600 (accessed 8 December 2023).

Yao, S., Chen, H., et al. (2023), ‘WebShop: Towards Scalable Real-World Web Interaction with Grounded Language Agents’. arXiv. Available at http://arxiv.org/abs/2207.01206 (accessed 11 December 2023).

Yao, S., Yu, D., et al. (2023), ‘Tree of Thoughts: Deliberate Problem Solving with Large Language Models’. arXiv. Available at http://arxiv.org/abs/2305.10601 (accessed 1 December 2023).

Yao, S., Zhao, J., et al. (2023), ‘ReAct: Synergizing Reasoning and Acting in Language Models’. arXiv. Available at http://arxiv.org/abs/2210.03629 (accessed 11 December 2023).

Ziegler, D. M. et al. (2019), ‘Fine-Tuning Language Models from Human Preferences’. arXiv. Available at https://doi.org/10.48550/arXiv.1909.08593.

Предыдущая главаГлава 49 из 52Следующая глава