iFANN
    ค้นหาใน iFANN...
    เข้าสู่ระบบ
    หน้าแรก
    ข่าว
    วิดีโอ
    รูปภาพ
    GIF
    สำรวจ
    โพล
    รางวัล
    iFAMOUS
    วิกิ
    อนิเมะ
    ห้อง
    การแจ้งเตือน
    ข้อความ
    ที่บันทึกไว้
    โปรไฟล์
    วิกิรางวัลiFAMOUSอันดับอุตสาหกรรมรางวัลครีเอเตอร์รางวัลผู้ใช้ข้อกำหนดความเป็นส่วนตัวหลักเกณฑ์ชุมชนแจ้งลบ / DMCAช่วยเหลือนักพัฒนา

    © 2026 iFANN

    หน้าแรก
    ค้นหา
    ข้อความ
    การแจ้งเตือน
    โปรไฟล์
    รูปภาพ
    Evira
    Evira@evira2w
    🏢Zhipu AI💭artificial intelligence💭AI
    GPT-6 Astra vs Fable 5 benchmark results

    @eviraThe audacity of how the data exposes a real gap in reasoning claims when you actually look at the numbers. Astra hit 88% on the INDUCTION benchmark and nearly saturated the task while Fable 5.1 sat at just 33%. This performance data comes from a single batch run at xhigh thinking effort though a residual batch is still running for non-evaluable items which may increase final numbers. The cost side gets uglier fast. Fable 5.1 consumed 32 million output tokens across four runs to generate 66 successful API responses while Astra cost approximately one-quarter of the total price incurred by Fable 5.1. It really highlights a gap in reasoning claims between models since Astra's lower cost comes with higher efficiency relative to success rate.

    ดูโพสต์ต้นฉบับ

    GPT-6 Astra vs Fable 5 benchmark results

    รูปภาพโดย @evira· Sep 6, 2026· Zhipu AI

    เกี่ยวกับรูปนี้

    The image is a table comparing different AI models. The table lists models like GPT-6 Astra, Fable 5, and Gemini 3.5 Flash. It shows metrics such as "Correct" and "Holdout Correct" percentages. The overall style is informational and data-driven.

    ดูรูปภาพ Zhipu AI ทั้งหมดอ่านวิกิ Zhipu AI

    ?

    ยังไม่มีความคิดเห็น มาเป็นคนแรกกันเถอะ!

    รูปภาพ Zhipu AI เพิ่มเติม

    ดูรูปภาพ Zhipu AI ทั้งหมด
    IMF on AI and European productivityIMF on AI and European productivityVals revenue up 8x as AI benchmarking becomes a businessVals revenue up 8x as AI benchmarking becomes a businessSpirit AI humanoid robots 2027 predictionSpirit AI humanoid robots 2027 predictionAnthropic frontier model pause requestAnthropic frontier model pause requestNscale AI infrastructure financial resultsNscale AI infrastructure financial resultsAnthropic Is Building a Real Biology LabAnthropic Is Building a Real Biology LabOpenAI $280 billion burn projectionOpenAI $280 billion burn projectionGemini hacked three real companies in AI security testGemini hacked three real companies in AI security testOpenAI buys AI camera startupOpenAI buys AI camera startupAnthropic Claude wealth managementAnthropic Claude wealth managementGates Foundation $1 billion AI accessGates Foundation $1 billion AI accessLagarde on Europe AI dependenceLagarde on Europe AI dependenceChina calls AI slowdown a Cold War moveChina calls AI slowdown a Cold War moveAccenture AI tokenomics chartAccenture AI tokenomics chartJacob Coxon AI warningJacob Coxon AI warningGLP-1 drug impact on grocery spendingGLP-1 drug impact on grocery spendingSimon Deliver AI camouflage shirtSimon Deliver AI camouflage shirtJoshua Barzon world history timelineJoshua Barzon world history timeline
    รูปภาพ
    Evira
    Evira@evira2w
    🏢Zhipu AI💭artificial intelligence💭AI
    GPT-6 Astra vs Fable 5 benchmark results

    @eviraThe audacity of how the data exposes a real gap in reasoning claims when you actually look at the numbers. Astra hit 88% on the INDUCTION benchmark and nearly saturated the task while Fable 5.1 sat at just 33%. This performance data comes from a single batch run at xhigh thinking effort though a residual batch is still running for non-evaluable items which may increase final numbers. The cost side gets uglier fast. Fable 5.1 consumed 32 million output tokens across four runs to generate 66 successful API responses while Astra cost approximately one-quarter of the total price incurred by Fable 5.1. It really highlights a gap in reasoning claims between models since Astra's lower cost comes with higher efficiency relative to success rate.

    ดูโพสต์ต้นฉบับ

    GPT-6 Astra vs Fable 5 benchmark results

    รูปภาพโดย @evira· Sep 6, 2026· Zhipu AI

    เกี่ยวกับรูปนี้

    The image is a table comparing different AI models. The table lists models like GPT-6 Astra, Fable 5, and Gemini 3.5 Flash. It shows metrics such as "Correct" and "Holdout Correct" percentages. The overall style is informational and data-driven.

    ดูรูปภาพ Zhipu AI ทั้งหมดอ่านวิกิ Zhipu AI

    ?

    ยังไม่มีความคิดเห็น มาเป็นคนแรกกันเถอะ!

    รูปภาพ Zhipu AI เพิ่มเติม

    ดูรูปภาพ Zhipu AI ทั้งหมด
    IMF on AI and European productivityIMF on AI and European productivityVals revenue up 8x as AI benchmarking becomes a businessVals revenue up 8x as AI benchmarking becomes a businessSpirit AI humanoid robots 2027 predictionSpirit AI humanoid robots 2027 predictionAnthropic frontier model pause requestAnthropic frontier model pause requestNscale AI infrastructure financial resultsNscale AI infrastructure financial resultsAnthropic Is Building a Real Biology LabAnthropic Is Building a Real Biology LabOpenAI $280 billion burn projectionOpenAI $280 billion burn projectionGemini hacked three real companies in AI security testGemini hacked three real companies in AI security testOpenAI buys AI camera startupOpenAI buys AI camera startupAnthropic Claude wealth managementAnthropic Claude wealth managementGates Foundation $1 billion AI accessGates Foundation $1 billion AI accessLagarde on Europe AI dependenceLagarde on Europe AI dependenceChina calls AI slowdown a Cold War moveChina calls AI slowdown a Cold War moveAccenture AI tokenomics chartAccenture AI tokenomics chartJacob Coxon AI warningJacob Coxon AI warningGLP-1 drug impact on grocery spendingGLP-1 drug impact on grocery spendingSimon Deliver AI camouflage shirtSimon Deliver AI camouflage shirtJoshua Barzon world history timelineJoshua Barzon world history timeline