www.science.org/toc/science/...
www.science.org/toc/science/...
Anyway, this is just a curiosity - language extrapolation is not an exact thought analogue - but it would be interesting to compare of different models address these sorts of questions. And interesting to see how their internal models do as well.
Anyway, this is just a curiosity - language extrapolation is not an exact thought analogue - but it would be interesting to compare of different models address these sorts of questions. And interesting to see how their internal models do as well.
Mais c'est loin d'être le cas, parce que les contrôles ciblent les contribuables et les dispositifs les plus
Mais c'est loin d'être le cas, parce que les contrôles ciblent les contribuables et les dispositifs les plus
I don’t know how much is data and how much is extrapolation.
I don’t know how much is data and how much is extrapolation.
...My grounding is in biology. You have no way to transmit pain, you have no pain. I find their definition of "pain" as unbelievable.
*The models are meant to model human behavior.*
There's no added function, it's extrapolation from a mathematical model.
...My grounding is in biology. You have no way to transmit pain, you have no pain. I find their definition of "pain" as unbelievable.
*The models are meant to model human behavior.*
There's no added function, it's extrapolation from a mathematical model.
一般化された変種により、学習者は出力空間における暗黙の報酬を推論することで、教師の能力を超えることができる。
しかし、言語モデルのヘッドはこの変化を異方的に減衰させる。つまり、教師の隠れ状態にエンコードされた変化の大部分は、その重みのごく一部としてロジットに伝達され、出力空間の外挿が依存するサンプリングされたトークンの対数確率比にはノイズが混入し、そのノイズが外挿によって増幅されることで、学習が不安定になる。
強化学習(RL)によ...
一般化された変種により、学習者は出力空間における暗黙の報酬を推論することで、教師の能力を超えることができる。
しかし、言語モデルのヘッドはこの変化を異方的に減衰させる。つまり、教師の隠れ状態にエンコードされた変化の大部分は、その重みのごく一部としてロジットに伝達され、出力空間の外挿が依存するサンプリングされたトークンの対数確率比にはノイズが混入し、そのノイズが外挿によって増幅されることで、学習が不安定になる。
強化学習(RL)によ...
My personal terminology for the weights was an n-dimensional calibration curve. Enough examples of arithmetic and you can make really good correlation output
My personal terminology for the weights was an n-dimensional calibration curve. Enough examples of arithmetic and you can make really good correlation output
But Iain Banks did that by authorial fiat. The Culture books aren't extrapolation, and longing is not going to make Claude grow up to be the Sleeper Service, or even Flere-Imsaho.
But Iain Banks did that by authorial fiat. The Culture books aren't extrapolation, and longing is not going to make Claude grow up to be the Sleeper Service, or even Flere-Imsaho.
If you legit can't tell whether it's AI because the passage is short or a human writer just kinda sucks, you can always just...hit the back button? Put it down? Without evidence or confession, it's just a witch hunt. And we
If you legit can't tell whether it's AI because the passage is short or a human writer just kinda sucks, you can always just...hit the back button? Put it down? Without evidence or confession, it's just a witch hunt. And we
(Yes, I know thats wild extrapolation from the original post and article)
@growing_daniel
My favorite part of this article is one of the visiting scholars pointing out that by Olah’s logic he is by far the worst slaver in history and he must immediately give up this abhorrent and immensely lucrative project, which anthropic people ignored.
(Yes, I know thats wild extrapolation from the original post and article)