• Home
  • Technology
  • Gaming
  • Entertainment
  • World & Business
  • Science
  • Sports
  • AI
HomeTechnologyGamingEntertainmentWorld & BusinessScienceSportsAI
  • HomeTechnologyGamingEntertainmentWorld & BusinessScienceSportsAI
    • Home
    • Technology
    • Gaming
    • Entertainment
    • World & Business
    • Science
    • Sports
    • AI
    AI

    Arena CEO Notes User Praise Gains for Gemini 3.8 Flash

    Arena co-founder highlights user satisfaction gains in Gemini 3.8 flash over prior version.

    AR
    JS
    AN
    6 Sources, 28d ago, first seen 28d ago

    TLDR

    Anastasios Nikolas Angelopoulos, co-founder and CEO of Arena, posted that Gemini 3.8 flash improves on 3.7 with a double-digit lift in explicit user praise. He also cited gains in steerability and bash. The post observes that models can still deliver raw 20 percent performance jumps. Jiao Sun, a Google DeepMind research engineer, retweeted the statement. The remarks come from public posts on X and reflect the author's stated observations about the model.

    Combined views

    273.1K

    6 Sources, first seen 28d ago

    Combined views

    273.1K

    6 Sources, first seen 28d ago

    1.8K likes
    1.8K likes
    90 comments
    222 saves
    128 reposts

    Sentiment

    Positive——Negative

    Summary

    Not enough discussion yet.

    No sentiment analysis available yet.

    90 comments
    222 saves
    128 reposts

    Sentiment

    Positive——Negative

    Summary

    Not enough discussion yet.

    No sentiment analysis available yet.

    Today's Rank

    —

    Not ranked yet

    Today's Rank

    —

    Not ranked yet

    6 Sources

    @arenaGemini 3.8 Flash (High) by @GoogleDeepMind is here! It just debuted across Agent Arena, Text Arena, and Code Arena: WebDev. In Agent Arena, it landed #14 with +5.94% net improvement. This ranks just above DeepSeek-V4-Pro at #15 (+5.91%) and is a significant jump from Gemini 3.7 Flash (High) at #32 (+0.84%). Its strongest signals are: +14.78% in Praise vs. Complaint (implicit sentiment in users reactions) +11.12% in Confirmed Success (explicit “yes, that worked” from users) In Text Arena, Gemini 3.8 Flash (High) is #7 with 1494 pts, ahead of Claude Opus 5 (High) at #8 (1492 pts) and Gemini 3.7 Flash (High) at #10 (1491 pts)! This release improved over Gemini 3.7 Flash (High) by category as well: - Writing, Literature & Language: #7 → #3 - Multi-Turn: #9 → #4 - Longer Query: #13 → #5 - Hard Prompts (English): #23 → #6 - Business, Management & Financial Ops: #26 → #7 - Coding: #22 → #7 - Hard Prompts: #13 → #7 - Instruction Following: #11 → #9 - Software & IT Services: #18 → #12 In the Code Arena: WebDev, Gemini 3.8 Flash (High) is #18 with 1567 points, enough to keep Google DeepMind at #8 when labs are ranked by their best-performing model. Congrats to the @GoogleDeepMind team on the release!
    @ml_angelopoulosGemini 3.8 flash is a big improvement over 3.7, especially in terms of user satisfaction. Double-digit lift in the percentage of explicit user praise. Also better at steerability and bash. It's cool to see that models can still see a raw 20% jump in performance these days.
    @sunjiao123sun_RT @ml_angelopoulos: Gemini 3.8 flash is a big improvement over 3.7, especially in terms of user satisfaction. Double-digit lift in the p…

    6 Sources

    @arenaGemini 3.8 Flash (High) by @GoogleDeepMind is here! It just debuted across Agent Arena, Text Arena, and Code Arena: WebDev. In Agent Arena, it landed #14 with +5.94% net improvement. This ranks just above DeepSeek-V4-Pro at #15 (+5.91%) and is a significant jump from Gemini 3.7 Flash (High) at #32 (+0.84%). Its strongest signals are: +14.78% in Praise vs. Complaint (implicit sentiment in users reactions) +11.12% in Confirmed Success (explicit “yes, that worked” from users) In Text Arena, Gemini 3.8 Flash (High) is #7 with 1494 pts, ahead of Claude Opus 5 (High) at #8 (1492 pts) and Gemini 3.7 Flash (High) at #10 (1491 pts)! This release improved over Gemini 3.7 Flash (High) by category as well: - Writing, Literature & Language: #7 → #3 - Multi-Turn: #9 → #4 - Longer Query: #13 → #5 - Hard Prompts (English): #23 → #6 - Business, Management & Financial Ops: #26 → #7 - Coding: #22 → #7 - Hard Prompts: #13 → #7 - Instruction Following: #11 → #9 - Software & IT Services: #18 → #12 In the Code Arena: WebDev, Gemini 3.8 Flash (High) is #18 with 1567 points, enough to keep Google DeepMind at #8 when labs are ranked by their best-performing model. Congrats to the @GoogleDeepMind team on the release!
    @ml_angelopoulosGemini 3.8 flash is a big improvement over 3.7, especially in terms of user satisfaction. Double-digit lift in the percentage of explicit user praise. Also better at steerability and bash. It's cool to see that models can still see a raw 20% jump in performance these days.
    @sunjiao123sun_RT @ml_angelopoulos: Gemini 3.8 flash is a big improvement over 3.7, especially in terms of user satisfaction. Double-digit lift in the p…