Комментарий
«А вон тот компьютер 2000 года выпуска, который ты сейчас выкинешь?», — только если он надоел в качестве подставки: один черт кроме этого он только шуметь и место занимать способен.
← О текущем моменте и смысле жизни
Читать и комментировать в ЖЖ ↗
«А вон тот компьютер 2000 года выпуска, который ты сейчас выкинешь?», — только если он надоел в качестве подставки: один черт кроме этого он только шуметь и место занимать способен.
С учетом того, что "успех" в минимизации свободной энергии (показатель "спасения") — это максимальное запутывание со средой, и в конечном счете, тепловая "смерть" вселенной (хотя, похоже, правильнее назвать тепловое "спасение") — это состояние максимальной запутанности всех степеней свободы во вселенной ("вся вселенная спаслась") — получается, что "смысл жизни — во спасении" — это переформулирование принципа унитарности. Это нефальсифицируемо (наверное, в рамках наших текущих представлений о физике), но и не очень интересно.
Интереснее, если мы принимаем этот принцип "спасения" — ответить на содержательный нормативный вопрос, как именно мы должны спасаться. FEP в текущем виде не дает ответа на следующие вопросы:
Поэтому работы по безмасштабной теории этики еще много. Правильнее даже сказать, что она еще толком не начиналась.
И еще вдогонку ссылка — надо ко всему этому добавить resource theories (https://ailev.livejournal.com/1567297.html). В quantum FEP статье говорится про разделение границы на "информационный" и "ресурсный/энергетический" сектора, но этого пока мало.
Вот грега игана люде нечиталле, а потом и удивляццо будут!..
Archive of chatGPT failures
https://docs.google.com/spreadsheets/d/1kDSERnROv5FgHbVN8z_bXH9gak2IXRtoqz0nwhrviCw/htmlview#gid=1302320625
Large language models have been demonstrated to be valuable in differentfields. ChatGPT, developed by OpenAI, has been trained using massive amounts ofdata and simulates human conversation by comprehending context and generatingappropriate responses. It has garnered significant attention due to its abilityto effectively answer a broad range of human inquiries, with fluent andcomprehensive answers surpassing prior public chatbots in both security andusefulness. However, a comprehensive analysis of ChatGPT's failures is lacking,which is the focus of this study. Ten categories of failures, includingreasoning, factual errors, math, coding, and bias, are presented and discussed.The risks, limitations, and societal implications of ChatGPT are alsohighlighted. The goal of this study is to assist researchers and developers inenhancing future language models and chatbots.
arxiv.org
This paper proposes a framework for quantitatively evaluating interactiveLLMs such as ChatGPT using publicly available data sets. We carry out anextensive technical evaluation of ChatGPT using 21 data sets covering 8different common NLP application tasks. We evaluate the multitask, multilingualand multi-modal aspects of ChatGPT based on these data sets and a newlydesigned multimodal dataset. We find that ChatGPT outperforms LLMs withzero-shot learning on most tasks and even outperforms fine-tuned models on sometasks. We find that it is better at understanding non-Latin script languagesthan generating them. It is able to generate multimodal content from textualprompts, via an intermediate code generation step. Moreover, we find thatChatGPT is 64.33% accurate on average in 10 different reasoning categoriesunder logical reasoning, non-textual reasoning, and commonsense reasoning,hence making it an unreliable reasoner. It is, for example, better at deductivethan inductive reasoning. ChatGPT suffers from hallucination problems likeother LLMs and it generates more extrinsic hallucinations from its parametricmemory as it does not have access to an external knowledge base. Finally, theinteractive feature of ChatGPT enables human collaboration with the underlyingLLM to improve its performance, i.e, 8% ROUGE-1 on summarization and 2% ChrF++on machine translation, in a multi-turn "prompt engineering" fashion.
arxiv.org
And I don't even start to put links to Gary Marcus' critics or the last revelations of Le Cun himself.