iFANN
    iFANN'da ara...
    Giriş Yap
    Ana Sayfa
    Haberler
    Videolar
    Fotoğraflar
    GIF'ler
    Keşfet
    Anketler
    Ödüller
    iFAMOUS
    Viki
    Anime
    Odalar
    Bildirimler
    Mesajlar
    Yer İmleri
    Profil
    VikiÖdülleriFAMOUSSıralamalarSektörlerİçerik Üretici ÖdülleriKullanıcı ÖdülleriŞartlarGizlilikTopluluk KurallarıKaldırma / DMCAYardımGeliştiriciler

    © 2026 iFANN

    Ana Sayfa
    Ara
    Mesajlar
    Uyarılar
    Profil
    Fotoğraf
    Estebankiwi
    Estebankiwi@estebankiwi2mo
    🏢Anthropic💭AI💭Tech
    Claude Sonnet 5 benchmarks vs Opus 4.8

    @estebankiwiAnthropic has released Claude Sonnet 5, priced at $3 per million input tokens and $15 per million output tokens,matching the cost of Sonnet 4.6. However, benchmark results show a significant leap: Sonnet 5 is now competitive with Opus 4.8 across all categories and even surpasses it on the GDPval-AA v2 knowledge benchmark. Details at and

    Orijinal gönderiyi gör

    Claude Sonnet 5 benchmarks vs Opus 4.8

    @estebankiwi tarafından fotoğraf· Jun 30, 2026· Anthropic

    Bu fotoğraf hakkında

    A clean comparison table highlights three AI models across five benchmark categories. Sonnet 5 occupies the left column with a light pink background and brown border, showing the highest scores in four of the five rows. Sonnet 4.6 sits in the middle column with lower percentages across the board. Opus 4.8 appears on the right as a reference model, posting the top marks in agentic coding and computer use. The table lists exact metrics like 63.2 percent for Sonnet 5 in agentic coding SWE-bench Pro and 1618 for Sonnet 5 in knowledge work GDPval-AA v2.

    Tüm Anthropic fotoğraflarını görAnthropic vikisini oku

    ?

    Daha fazla Anthropic fotoğrafı

    Tüm Anthropic fotoğraflarını gör
    BridgeMind Codex outageBridgeMind Codex outageGPT Astra vs Fable 5.1 benchmarkGPT Astra vs Fable 5.1 benchmarkQwen 3.8 27B ties GPT 5.6 LunaQwen 3.8 27B ties GPT 5.6 LunaCodex features and uptime chartCodex features and uptime chartGPT Astra ThinkingGPT Astra ThinkingQwen 3.8 Max TerminalBench scoreQwen 3.8 Max TerminalBench scoreDeepSeek v4 Flash 284B beats GPT 5.6 LunaDeepSeek v4 Flash 284B beats GPT 5.6 LunaClaude Opus 5 vs GPT 5.6 SolClaude Opus 5 vs GPT 5.6 SolKimi K3 sold out plansKimi K3 sold out plansKimi K3 vs Claude Opus on AI IndexKimi K3 vs Claude Opus on AI IndexKimi K3 vs Fable 5 game development pollKimi K3 vs Fable 5 game development pollKimi K3 launch rumorsKimi K3 launch rumorsGoogle Gemini 3.5 Pro release chartGoogle Gemini 3.5 Pro release chartOpenAI Codex unlimited usageOpenAI Codex unlimited usageArtificial Analysis Coding Index Fable 5Artificial Analysis Coding Index Fable 5Grok 4.5 CursorBench results cost comparisonGrok 4.5 CursorBench results cost comparisonGrok 4.5 live in CursorGrok 4.5 live in CursorBridgeMind Fable 5 return partyBridgeMind Fable 5 return party
    Fotoğraf
    Estebankiwi
    Estebankiwi@estebankiwi2mo
    🏢Anthropic💭AI💭Tech
    Claude Sonnet 5 benchmarks vs Opus 4.8

    @estebankiwiAnthropic has released Claude Sonnet 5, priced at $3 per million input tokens and $15 per million output tokens,matching the cost of Sonnet 4.6. However, benchmark results show a significant leap: Sonnet 5 is now competitive with Opus 4.8 across all categories and even surpasses it on the GDPval-AA v2 knowledge benchmark. Details at and

    Orijinal gönderiyi gör

    Claude Sonnet 5 benchmarks vs Opus 4.8

    @estebankiwi tarafından fotoğraf· Jun 30, 2026· Anthropic

    Bu fotoğraf hakkında

    A clean comparison table highlights three AI models across five benchmark categories. Sonnet 5 occupies the left column with a light pink background and brown border, showing the highest scores in four of the five rows. Sonnet 4.6 sits in the middle column with lower percentages across the board. Opus 4.8 appears on the right as a reference model, posting the top marks in agentic coding and computer use. The table lists exact metrics like 63.2 percent for Sonnet 5 in agentic coding SWE-bench Pro and 1618 for Sonnet 5 in knowledge work GDPval-AA v2.

    Tüm Anthropic fotoğraflarını görAnthropic vikisini oku

    ?

    Daha fazla Anthropic fotoğrafı

    Tüm Anthropic fotoğraflarını gör
    BridgeMind Codex outageBridgeMind Codex outageGPT Astra vs Fable 5.1 benchmarkGPT Astra vs Fable 5.1 benchmarkQwen 3.8 27B ties GPT 5.6 LunaQwen 3.8 27B ties GPT 5.6 LunaCodex features and uptime chartCodex features and uptime chartGPT Astra ThinkingGPT Astra ThinkingQwen 3.8 Max TerminalBench scoreQwen 3.8 Max TerminalBench scoreDeepSeek v4 Flash 284B beats GPT 5.6 LunaDeepSeek v4 Flash 284B beats GPT 5.6 LunaClaude Opus 5 vs GPT 5.6 SolClaude Opus 5 vs GPT 5.6 SolKimi K3 sold out plansKimi K3 sold out plansKimi K3 vs Claude Opus on AI IndexKimi K3 vs Claude Opus on AI IndexKimi K3 vs Fable 5 game development pollKimi K3 vs Fable 5 game development pollKimi K3 launch rumorsKimi K3 launch rumorsGoogle Gemini 3.5 Pro release chartGoogle Gemini 3.5 Pro release chartOpenAI Codex unlimited usageOpenAI Codex unlimited usageArtificial Analysis Coding Index Fable 5Artificial Analysis Coding Index Fable 5Grok 4.5 CursorBench results cost comparisonGrok 4.5 CursorBench results cost comparisonGrok 4.5 live in CursorGrok 4.5 live in CursorBridgeMind Fable 5 return partyBridgeMind Fable 5 return party