İçeriğe atla
0
  • Ders Ara
  • Ana Sayfa
  • Kategoriler
    • All Categories
      • Individual Categories
    • Gruplar
    • Okunmamış 0
    • Güncel
    • Kullanıcılar
    • Hakkımızda
    • Öğrenci Fırsatları
    • Akademik Takvim
    • CV oluşturucu
    • IEU Timetable
    • Devamsızlık App
    • IEU GPA Hesaplayıcı
    • Niki Cüzdan
    • Ders Ara
    • Ana Sayfa
    • Kategoriler
      • All Categories
        • Individual Categories
      • Gruplar
      • 0 Okunmamış 0
      • Güncel
      • Kullanıcılar
      • Hakkımızda
      • Öğrenci Fırsatları
      • Akademik Takvim
      • CV oluşturucu
      • IEU Timetable
      • Devamsızlık App
      • IEU GPA Hesaplayıcı
      • Niki Cüzdan
      Daralt
      IEU Forum – İzmir Ekonomi Üniversitesi Öğrenci Topluluğu Platformu

      IEU Forum

      -- çevrimiçi
      1. Ana Sayfa
      2. AI
      3. Are Latent Reasoning Models Easily Interpretable?

      Final Unicourse'tan Çalış, Yüksek Notu Garantile!

      %25 İndirim Kodu: FRM25
      Yükleniyor...
      Dersi İzle
      GÖRÜNTÜLEYENLER
      +36
      Premium Özellik
      Bu konuyu kimlerin görüntülediğini görmek için Premium üyelik gerekir.
      Premium'a Geç

      Vizesine Unicourse'tan Çalış, Yüksek Notu Garantile!

      A B C D Çıkmış Sorular Formül Kağıtları Konu Anlatımı Sınav İpuçları Örnek Sınav
      Dersi İzle
      YENİ ÖZELLİK

      Bi'Öğrenci Fırsatları
      Forum'da!

      Bi'Öğrenci ile artık forum üzerinden en güncel indirimlere, anlık fırsatlara ve avantajlı tekliflere ulaşabilirsin.

      FIRSATLARI KEŞFET
      Red Bull Basement
      SPONSORLU ETKİNLİK

      Fikrini Gerçeğe Dönüştür

      Projeni dünyaya göstermek için sahne hazır. Red Bull Basement başvuruları açık.

      Başvurunu Yap

      🎉 Foruma Yeni Özellik Geldi!

      Sizin için PDF toollarını getirdik!

      İncele ve Kullan

      Are Latent Reasoning Models Easily Interpretable?

      Konu Zamanlandı Sabitlendi Kilitli Taşındı AI
      artificialintel
      1 İleti 1 Yayımlayıcılar 0 Bakış
      • En eskiden en yeniye
      • En yeniden en eskiye
      • En çok oylanan
        Cevap
        • Yeni başlık oluşturarak cevapla
        Cevaplamak için giriş yapın
        Bu başlık silindi. Sadece başlık düzenleme yetkisi olan kullanıcılar görebilir.
        • yogthos@lemmy.mlY This user is from outside of this forum
          yogthos@lemmy.mlY This user is from outside of this forum
          yogthos@lemmy.ml
          yazdı Son düzenleyen:
          #1

          Models normally do all their reasoning in a continuous hidden state instead of spitting out readable text which makes them hard to monitor. The authors tested the Coconut and CODI models and it turns out these models barely even use their hidden reasoning steps for logical tasks like PrOntoQA and ProsQA. You can force the models to stop thinking early and they almost always spit out the same response anyway. It turns out that their high performance on logical tasks actually comes from their specific training data rather than the extra thinking during inference.

          Things get even more interesting when the models actually need those reasoning tokens for math problems. The researchers wanted to know if standard step-by-step math solutions were hidden inside the latent space, and projected the hidden states back into regular vocabulary words to check. And sure enough when the models got the math problem right the researchers found the correct intermediate math steps in their hidden states up to 93% of the time. The finding strongly suggests that the models are basically doing standard math steps in the background.

          They confirmed the exact math operations taking place by tweaking numbers in the prompt and seeing how the hidden states reacted which allowed decoding a verified reasoning path for a large majority of correct predictions. But they could rarely do this for incorrect predictions proving that models are actually way more interpretable than the AI community assumed. And you can even use that interpretability as a signal to guess if the model is about to give a right or wrong answer.

          https://arxiv.org/abs/2604.04902

          1 Cevap Son cevap
          1

          Hello! It looks like you're interested in this conversation, but you don't have an account yet.

          Getting fed up of having to scroll through the same posts each visit? When you register for an account, you'll always come back to exactly where you were before, and choose to be notified of new replies (either via email, or push notification). You'll also be able to save bookmarks and upvote posts to show your appreciation to other community members.

          With your input, this post could be even better 💗

          Kayıt Ol Giriş
          Cevap
          • Yeni başlık oluşturarak cevapla
          Cevaplamak için giriş yapın
          • En eskiden en yeniye
          • En yeniden en eskiye
          • En çok oylanan


            Önerilen Başlıklar

            • U

              South Korea’s ‘AI for All’ Push Gives Free Access to Every Citizen

              Takip ediliyor Susturulmuş Konu Zamanlandı Sabitlendi Kilitli Taşındı AI artificialintel
              1
              1 Oy
              1 İleti
              0 Bakış
              Kimse yanıtlamadı
            • yogthos@lemmy.mlY

              Canonical-basis realignment for Transformer LLMs: every hidden axis becomes independently measurable and controllable.

              Takip ediliyor Susturulmuş Konu Zamanlandı Sabitlendi Kilitli Taşındı AI artificialintel
              1
              1 Oy
              1 İleti
              0 Bakış
              Kimse yanıtlamadı
            • pete_link@lemmy.mlP

              'I've Never Seen This': Massive Collapse in Support for AI Data Centers Captured in New Poll | Common Dreams

              Takip ediliyor Susturulmuş Konu Zamanlandı Sabitlendi Kilitli Taşındı AI artificialintel
              1
              1 Oy
              1 İleti
              0 Bakış
              Kimse yanıtlamadı
            • yogthos@lemmy.mlY

              WorldClaw: Agentic 3D Open-World Generation at Scale

              Takip ediliyor Susturulmuş Konu Zamanlandı Sabitlendi Kilitli Taşındı AI artificialintel
              1
              1 Oy
              1 İleti
              0 Bakış
              Kimse yanıtlamadı
            • I

              Has anybody had good luck with OpenWebUI Computer?

              Takip ediliyor Susturulmuş Konu Zamanlandı Sabitlendi Kilitli Taşındı AI artificialintel
              1
              1 Oy
              1 İleti
              0 Bakış
              Kimse yanıtlamadı

            Developed by Enes Uysal & Kadir Ay

            9

            Çevrimiçi

            8.8k

            Kullanıcı

            1.9k

            Konu

            3.7k

            İleti
            • Giriş

            • Hesabınız yok mu? Kayıt Ol

            • Aramak için giriş yapın veya kaydolun
            • İlk ileti
              Son ileti