问 HN:基于 Transformer 的 LLM 会达到改进上限吗?
1 分•作者: jaguar75•大约 1 年前
我有一个根本性的问题:当前基于 Transformer 的架构是否足够强大,能够实现向通用人工智能(AGI)的扩展,或者说,是否存在越来越明显的上限?
虽然来自领先的大型语言模型(LLM)公司的专家们站在最前沿,但由于明显的利益冲突,很难完全相信他们的话。而且,目前尚不清楚学术界是否在这方面拥有更多的理论知识,因为这些商业公司似乎也在引领着相关研究工作。
查看原文
A fundamental question I have is whether the current transformer based architectures are powerful enough to allow scaling towards AGI or are there ceilings here that are becoming more evident?<p>Though folks from the leading LLM companies are in the forefront of this, it’s hard to take their word due to an obvious conflict of interest. And it’s unclear if academia has more theoretical knowledge here as it seems these commercial companies are also leading the research work here.