Seenthis
•
 
Identifiants personnels
  • [mot de passe oublié ?]

METR

https://metr.org

  • ►/blog
    • ►/2025-07-10-early-2025-ai-experienced-os-dev-study
  • @biggrizzly
    BigGrizzly @biggrizzly CC BY-NC-SA 14/01/2026
    3
    @rastapopoulos
    @alexcorp
    @ericw
    3

    If AI coding is so good … where are the performance numbers? – Pivot to AI
    ►https://pivot-to-ai.com/2026/01/13/if-ai-coding-is-so-good-where-are-the-performance-numbers

    (...)

    We have one public study of AI coding performance that applied any reasonable methodology to the coding itself. That’s the METR study from July last year.

    METR got 16 experienced open-source developers with “moderate AI experience.” The devs fixed real bug reports in their own projects. They used either Cursor with Claude Code or no AI help, at random.

    METR actually screen-recorded and timed the work. The devs said they’d worked 20% faster — but they’d actually been slowed down by 19%. [blog post; paper, PDF]
    ►https://metr.org/blog/2025-07-10-early-2025-ai-experienced-os-dev-study

    And that’s why you have to measure. Vibe self-reports are wrong. You must measure.

    (...)

    In fact, IEEE Spectrum ran a story where an ardent vibe coder notices exactly that: “AI Coding Assistants Are Getting Worse: Newer models are more prone to silent but deadly failure modes.” [IEEE]

    – A task that might have taken five hours assisted by AI, and perhaps 10 hours without it, is now more commonly taking seven or eight hours, or even longer. It’s reached the point where I am sometimes going back and using older versions of large language models.

    (...)

    En rapport aussi avec :
    ▻https://seenthis.net/messages/1153910

    AI Coding Degrades : Silent Failures Emerge - IEEE Spectrum
    ►https://spectrum.ieee.org/ai-coding-degrades

    BigGrizzly @biggrizzly CC BY-NC-SA
    • @ericw
      EricW @ericw CC BY-SA 15/01/2026

      #mais_quelle_surprise

      EricW @ericw CC BY-SA
    Écrire un commentaire
  • @biggrizzly
    BigGrizzly @biggrizzly CC BY-NC-SA 11/07/2025
    4
    @simplicissimus
    @aurelieng
    @b_b
    @cy_altern
    4

    sebsauvage - Mastodon
    ▻https://framapiaf.org/@sebsauvage/114833287629627467

    #IA #développement

    Les boîtes qui vendent l’IA prétendent que leur produit va vous faire gagner énormément de temps en développement.

    Une étude scientifique a été réalisée.
    Conclusion de cette étude : Les développeurs qui utilisent des IA mettent 19% de temps en PLUS à accomplir les tâches.

    Effet psychologique intéressant : Les développeurs utilisant des IA (donc 19% plus lents) étaient persuadés d’être 20% plus rapides que les autres.

    Measuring the Impact of Early-2025 AI on Experienced Open-Source Developer Productivity - METR
    ►https://metr.org/blog/2025-07-10-early-2025-ai-experienced-os-dev-study

    We conduct a randomized controlled trial (RCT) to understand how early-2025 AI tools affect the productivity of experienced open-source developers working on their own repositories. Surprisingly, we find that when developers use AI tools, they take 19% longer than without—AI makes them slower. We view this result as a snapshot of early-2025 AI capabilities in one relevant setting; as these systems continue to rapidly evolve, we plan on continuing to use this methodology to help estimate AI acceleration from AI R&D automation .

    BigGrizzly @biggrizzly CC BY-NC-SA
    • @b_b
      b_b @b_b PUBLIC DOMAIN 17/07/2025

      Cité ici aussi ▻https://next.ink/192648/la-productivite-des-developpeurs-semble-baisser-quand-ils-utilisent-lia-genera

      b_b @b_b PUBLIC DOMAIN
    Écrire un commentaire