เธรด

    โพสต์ต้นฉบับถูกลบไปแล้ว
    Astrid

    The models always used internal reasoning; predicting the next token was simply the output method. Some people experimented to see if the models could forecast 2 or 3 tokens ahead in the output, and they can.

    ที่แล้ว

    0 ถูกใจ0 ไม่ถูกใจ0 การตอบกลับ
    ?

    ยังไม่มีการตอบกลับ