iFANN
    iFANN 검색...
    로그인
    홈
    뉴스
    동영상
    사진
    GIF
    탐색
    투표
    어워드
    iFAMOUS
    위키
    애니
    룸
    알림
    메시지
    북마크
    프로필
    위키어워드iFAMOUS랭킹산업크리에이터 리워드사용자 리워드약관개인정보커뮤니티 가이드라인게시 중단 / DMCA도움말개발자

    © 2026 iFANN

    홈
    검색
    메시지
    알림
    프로필
    사진
    Evira
    Evira@evira2w
    🏢Zhipu AI💭artificial intelligence💭AI
    GPT-6 Astra vs Fable 5 benchmark results

    @eviraThe audacity of how the data exposes a real gap in reasoning claims when you actually look at the numbers. Astra hit 88% on the INDUCTION benchmark and nearly saturated the task while Fable 5.1 sat at just 33%. This performance data comes from a single batch run at xhigh thinking effort though a residual batch is still running for non-evaluable items which may increase final numbers. The cost side gets uglier fast. Fable 5.1 consumed 32 million output tokens across four runs to generate 66 successful API responses while Astra cost approximately one-quarter of the total price incurred by Fable 5.1. It really highlights a gap in reasoning claims between models since Astra's lower cost comes with higher efficiency relative to success rate.

    원본 게시물 보기

    GPT-6 Astra vs Fable 5 benchmark results

    @evira님의 사진· Sep 6, 2026· Zhipu AI

    이 사진에 대해

    The image is a table comparing different AI models. The table lists models like GPT-6 Astra, Fable 5, and Gemini 3.5 Flash. It shows metrics such as "Correct" and "Holdout Correct" percentages. The overall style is informational and data-driven.

    Zhipu AI 사진 전체 보기Zhipu AI 위키 읽기

    ?

    아직 댓글이 없습니다. 첫 댓글을 남겨보세요!

    Zhipu AI 사진 더 보기

    Zhipu AI 사진 전체 보기
    EU policy on children under 152EU policy on children under 15AI models Kimi Qwen DeepSeek GLM MiniMaxAI models Kimi Qwen DeepSeek GLM MiniMaxGPT Astra vs Fable 5.1 benchmarkGPT Astra vs Fable 5.1 benchmarkClaude vs Chinese model history promptClaude vs Chinese model history promptAI agent security risk alert2AI agent security risk alertSenseNova U1 Pro open-source modelSenseNova U1 Pro open-source modelGLM-5.2 pricing vs OpenRouter2GLM-5.2 pricing vs OpenRouterZ.ai GLM 5.2 OpenRouter pricing 1M contextZ.ai GLM 5.2 OpenRouter pricing 1M contextGLM rate limit errorGLM rate limit error
    사진
    Evira
    Evira@evira2w
    🏢Zhipu AI💭artificial intelligence💭AI
    GPT-6 Astra vs Fable 5 benchmark results

    @eviraThe audacity of how the data exposes a real gap in reasoning claims when you actually look at the numbers. Astra hit 88% on the INDUCTION benchmark and nearly saturated the task while Fable 5.1 sat at just 33%. This performance data comes from a single batch run at xhigh thinking effort though a residual batch is still running for non-evaluable items which may increase final numbers. The cost side gets uglier fast. Fable 5.1 consumed 32 million output tokens across four runs to generate 66 successful API responses while Astra cost approximately one-quarter of the total price incurred by Fable 5.1. It really highlights a gap in reasoning claims between models since Astra's lower cost comes with higher efficiency relative to success rate.

    원본 게시물 보기

    GPT-6 Astra vs Fable 5 benchmark results

    @evira님의 사진· Sep 6, 2026· Zhipu AI

    이 사진에 대해

    The image is a table comparing different AI models. The table lists models like GPT-6 Astra, Fable 5, and Gemini 3.5 Flash. It shows metrics such as "Correct" and "Holdout Correct" percentages. The overall style is informational and data-driven.

    Zhipu AI 사진 전체 보기Zhipu AI 위키 읽기

    ?

    아직 댓글이 없습니다. 첫 댓글을 남겨보세요!

    Zhipu AI 사진 더 보기

    Zhipu AI 사진 전체 보기
    EU policy on children under 152EU policy on children under 15AI models Kimi Qwen DeepSeek GLM MiniMaxAI models Kimi Qwen DeepSeek GLM MiniMaxGPT Astra vs Fable 5.1 benchmarkGPT Astra vs Fable 5.1 benchmarkClaude vs Chinese model history promptClaude vs Chinese model history promptAI agent security risk alert2AI agent security risk alertSenseNova U1 Pro open-source modelSenseNova U1 Pro open-source modelGLM-5.2 pricing vs OpenRouter2GLM-5.2 pricing vs OpenRouterZ.ai GLM 5.2 OpenRouter pricing 1M contextZ.ai GLM 5.2 OpenRouter pricing 1M contextGLM rate limit errorGLM rate limit error