AINewsnow

¿Puede una función de activación entrenable mejorar un sistema de reconocimiento de voz? La respuesta es sí, y los números lo confirman?

Recientemente realicé una prueba con datos reales de audio usando el dataset público Google Speech Commands (4 clases: yes, no, up, down; ~800 muestras; MFCC de 20 coeficientes; CNN ligera). El objetivo era simple: comparar mi tecnología Genal Activation Family contra las funciones estándar de la i…

Read the full story at r/learnmachinelearning ↗

Timeline · 2 reports

  1. 2026-10-06 14:36 · r/deeplearning
    ¿Puede una función de activación entrenable mejorar un sistema de reconocimiento de voz? La respuesta es sí, y los números lo confirman?
  2. 2026-10-06 14:36 · r/learnmachinelearning
    ¿Puede una función de activación entrenable mejorar un sistema de reconocimiento de voz? La respuesta es sí, y los números lo confirman?

More stories

  1. Google DeepMind launches EmbeddingGemma 2, a 740M-parameter model to map code, images, video, and audio in a shared embedding space, under an Apache 2.0 license (Google) — Techmeme
  2. Nano Banana 2.1 is rolling out now. — r/GeminiAI
  3. Google is about to remove free access to Gemini Flash and Pro — The Verge AI
  4. OpenAI is adding text watermarking in ChatGPT and Codex — The Verge AI
  5. I noticed this happens to nearly all Google models even Nanobana on web now is nerfed — r/GeminiAI
  6. OpenAI deactivated my account for "Cyber Abuse" while I was building a remote ADB support tool. Appeal rejected with no explanation. — r/OpenAI
  7. Create your own voices with Gemini 3.8 text-to-speech — Google DeepMind YouTube
  8. Day 4 no Gemini 4 — r/GeminiAI

Get the daily brief of stories like this at 6:30 every morning →